
Microsoft Bets on AMD's Helios to Crack Nvidia's Grip
Microsoft will deploy AMD's new Helios AI racks across Azure, joining Meta, Oracle and OpenAI as flagship customers in a direct challenge to Nvidia's 95% grip on the data center GPU market.
They summarize our coverage. We write it.
Newsletters like this one rebroadcast our headlines - often without the full review, the source reading, or the analysis underneath. Our weekly briefing sends the work they paraphrase, straight from the desk, before they get to it.
Free, weekly, no spam. One email every Tuesday. Unsubscribe anytime.

Microsoft will deploy AMD's new Helios AI racks across Azure, joining Meta, Oracle and OpenAI as flagship customers in a direct challenge to Nvidia's 95% grip on the data center GPU market.

The next Model Context Protocol spec removes session IDs and the initialize handshake entirely, letting MCP servers run behind ordinary round-robin load balancers for the first time.

Google is reportedly building a chip line separate from its TPUs that hardwires parts of Gemini directly into silicon, promising up to 10x efficiency as a capacity crunch forces Cloud to turn away customers.

SK Group Chairman Chey Tae-won says customers want 60-100% more AI memory in 2027 than in 2026, and warns governments are starting to treat chip access as a matter of economic security.

Internet pioneer Vint Cerf has joined Innovation Labs to push DNSid, a DNS-anchored identity standard for AI agents, through the IETF after retiring from Google.

Governor Hochul signed an executive order pausing permits for data centers over 50 megawatts for up to a year, making New York the first US state to enact a statewide moratorium.

Nebius will sell Reflection AI over $1 billion in Nvidia GB300 compute through 2029, the open-source AI lab's second billion-dollar infrastructure deal in three weeks.

Nvidia has joined Gradium's seed round while holding a stake in rival ElevenLabs, tying its GPUs to both leaders in the race to own voice AI.

GPT-5.6 Sol is now live on Cerebras wafer-scale hardware at 750 tokens per second - roughly 10x faster than any GPU-based frontier model deployment in production.

A Reuters-obtained internal memo reveals Meta will start manufacturing its Iris chip in September, targeting 14 GW of computing capacity by 2027 while cutting NVIDIA and AMD dependency.

SambaNova closes a $1B Series F at $11B valuation with JPMorgan Chase as its flagship on-premises inference partner and SN50 chips due in H2 2026.

ZML's LLMD inference server runs Llama, Qwen, and Mistral on Nvidia, AMD, TPU, Apple Metal, and Intel Arc from a single binary - for free.