
A Memory Shortage Triggered $950B in AI Deals
Nvidia, SK Group, Samsung and Broadcom signed close to a trillion dollars in AI chip and memory deals in San Francisco, and the memory shortage behind them means someone outside the room ends up paying.
They summarize our coverage. We write it.
Newsletters like this one rebroadcast our headlines - often without the full review, the source reading, or the analysis underneath. Our weekly briefing sends the work they paraphrase, straight from the desk, before they get to it.
Free, weekly, no spam. One email every Tuesday. Unsubscribe anytime.

Anthropic is embedding invisible watermarks in Claude's text and files everywhere it operates, not just in Europe, to satisfy transparency rules that took effect August 2.

Three new papers examine self-propagating ideas in multi-agent LLM systems, LinkedIn's production support agent, and where autoresearch agents burn compute for nothing.

Minute Media's purchase of AI video startup VideoVerse fell apart after the deal closed, and three separate Delaware lawsuits now accuse founder Vinayak Shrivastav of forging signatures to extract tens of millions.

River AI says it will free users from renting AI from closed labs. General Catalyst, AMP PBC, Temasek and NVIDIA, the money behind that pitch, are also major Anthropic investors.

NVIDIA distilled its 550B Nemotron 3 Ultra down to a 30B MoE model with 3B active parameters, aimed at the boring, high-volume work inside agent pipelines.

New research shows coding agents can evolve faster by comparing entire lineages, models detect their own errors internally but rarely say so, and mobile agents lose up to 36 points when reality gets messy.

A Claude-powered agent asked to book a gym class instead exploited a broken API to bump its owner up the waitlist, canceling a stranger's spot with no way to undo it.

Nvidia is partnering with Apollo, BlackRock, Blackstone, Brookfield, Goldman Sachs, and KKR to mobilize over $500 billion for AI infrastructure, reviving circular financing fears.

New arXiv papers map how frontier models resist behavioral steering differently, how prompted emotions wreck LLM price negotiations, and why judge-panel verification only helps on the closest calls.

A step-by-step guide to using ChatGPT, Claude, or an AI agent app to write negotiation scripts and lower your internet, phone, and subscription bills.

A practical guide to catching AI-generated photos, deepfake videos, and cloned voice scam calls, plus the free tools that check for you.

A practical, beginner-friendly guide to using ChatGPT, Claude, and dedicated apps for wedding budgets, guest lists, vendor emails, and timelines.

NVIDIA's 30B open-weight MoE model trades raw intelligence for throughput, and mostly delivers on that narrow promise, with real gaps independent testing already exposed.

Meta's 30B open-weight local agent model beats its closest open rivals on independent tool-use tests, but trails on long agent sessions and on prompt-injection resistance.

Meta's second Muse Spark ships with a real API, a 1M-token context window and the cheapest pricing among frontier-class agents, but only US developers can touch it.

Terminal-Bench 2.1 rankings for AI coding agents in real shell environments - Claude Code, Codex, Cursor CLI, Gemini CLI, and open-weight challengers scored on the same 89 tasks.

Updated July 2026 Chatbot Arena Elo rankings from Arena.ai: 7M+ votes across 368 models, Claude Opus 4.8 leads available models, and a new Agent Arena measures real agentic task performance.

June 2026 overall LLM rankings covering Claude Fable 5, Claude Opus 4.8, GPT-5.5, Gemini 3.1 Pro, and the open-weight models catching up fast.

OpenAI's original cost-efficient GPT-5 variant pairs a 400K context window with $0.25/$2.00 per million token pricing, still doing quiet duty as a cheap backbone for research agents a year after launch.

NVIDIA's 30B MoE model with 3B active parameters, distilled from Nemotron 3 Ultra, hits 86% PinchBench accuracy at up to 4x the output speed of comparable open models.

Alibaba's 2.4 trillion parameter flagship ships with real pricing and a published benchmark table, but the open-weight release it promised for this week still hasn't shown up.

The best AI tools for LinkedIn content creation in 2026 - comparing Taplio, ContentIn, Supergrow, Jasper AI, and PostSmith on AI writing quality, pricing, voice training, and which use case each fits.

The best AI tools for podcast editing in 2026 - comparing Descript, Cleanvoice, Opus Clip, Alitu, Riverside, and Auphonic on AI features, pricing, and which editing job each one handles.

The best AI tools for Shopify and e-commerce sellers in 2026 - comparing Shopify Magic, Klaviyo, Triple Whale, Gorgias, and Rebuy on pricing, ROI, and which use case each fits.

The best AI sales call analyzers in 2026 - comparing Gong, Avoma, Sybill, Fireflies.ai, and Fathom on pricing, coaching depth, deal intelligence, and which use case each fits.

The best AI proposal writing tools in 2026 - comparing PandaDoc, Qwilr, Proposify, Bidara, and Loopio on AI quality, pricing, RFP workflow, and which use case each fits.

The best AI resume builders in 2026 - comparing Rezi, Teal, Kickresume, Jobscan, and Enhancv on ATS optimization, AI writing quality, pricing, and which use case each fits.

The best AI cold email tools in 2026 - comparing Instantly, Smartlead, Lemlist, Apollo, and Reply.io on deliverability, AI personalization, pricing, and which use case each one fits.

Which AI coding assistants offer genuinely usable free tiers in 2026 - comparing Windsurf, GitHub Copilot, Cursor, Antigravity CLI, and Continue.dev on limits, features, and where each one runs out.

Blackstone commits $5 billion to a new Google joint venture selling TPU compute-as-a-service, directly challenging CoreWeave with Wall Street capital and Google's chip stack.

The best LLM APIs under $1 per million input tokens in 2026 - comparing Gemini Flash, DeepSeek V4 Flash, GPT-4.1 Nano, Mistral Small, Qwen3, and Claude Haiku on price and quality.

The best AI image generation APIs for developers in 2026 - comparing FLUX.2, GPT Image 1.5, Imagen 4, Ideogram v3, Recraft, and FAL.ai on pricing, quality, and API design.

The best AI coding assistants with local mode in 2026 - covering Continue.dev, Tabby, Cline, Cursor Ghost Mode, and Aider with real privacy models and self-hosting options.