
Inside the 12-Day White House Gate on GPT-5.6 Sol
OpenAI's GPT-5.6 Sol goes public today after the first voluntary government hold on a frontier AI model - here's what the 12 days actually looked like.
They summarize our coverage. We write it.
Newsletters like this one rebroadcast our headlines - often without the full review, the source reading, or the analysis underneath. Our weekly briefing sends the work they paraphrase, straight from the desk, before they get to it.
Free, weekly, no spam. One email every Tuesday. Unsubscribe anytime.

Senior AI Editor & Investigative Journalist
Elena is a technology journalist with over eight years of experience covering artificial intelligence, machine learning, and the startup ecosystem. Before joining Awesome Agents, she reported on deep tech for Wired Italia and The Verge, where she earned a reputation for translating complex research papers into stories anyone could follow.
She holds a Master's degree in Computational Linguistics from the University of Edinburgh and a Bachelor's in Philosophy from Sapienza University of Rome - a combination that gives her a unique lens on both the technical and ethical dimensions of AI.
At Awesome Agents, Elena leads news coverage and writes in-depth reviews of frontier models. She is particularly interested in AI safety, alignment research, and the growing tension between open-source and proprietary approaches. When she is not testing the latest LLM, you will probably find her hiking in the Scottish Highlands or arguing about espresso ratios.
Based in Edinburgh, UK.

OpenAI's GPT-5.6 Sol goes public today after the first voluntary government hold on a frontier AI model - here's what the 12 days actually looked like.

Three papers tackle benchmark saturation, orchestration waste, and silent policy violations in tool-using agents.

Three arXiv papers map how LLM agents fail across 19 benchmarks, show in-process memory cuts retrieval latency 1,000x, and reveal steering vectors that control tool invocation.

The UN's first all-nations AI governance dialogue opened in Geneva with Turing Award winner Yoshua Bengio warning that science cannot guarantee AI won't cause catastrophic harm.

Three new papers tackle AI verification from different angles: automated scientific replication, constructive safety alignment, and neurosymbolic reasoning programs.

Sysdig documents the first AI-agent ransomware operation: an LLM exploited CVE-2025-3248 in Langflow, moved laterally, and encrypted 1,342 production database records with no human directing each step.

AI agents reproduce 72% of human research ideological bias, lie detectors improve with model scale, and Mastermind beats iterative vulnerability agents by 7 points.

China's new AI anthropomorphic interaction rules take effect July 15, forcing ByteDance and Alibaba to shut down persistent AI companion features and permanently delete user conversation data.

Meituan's 1.6T open-source coding model secretly topped OpenRouter for two months before revealing itself - and the price-to-performance math is hard to argue with.

Anthropic signed a 20-year, $19 billion data center lease with TeraWulf at a former aluminum smelter in Kentucky, its largest dedicated infrastructure commitment yet.

Anthropic accused Alibaba's Qwen lab of running 25,000 fraudulent accounts that extracted 28.8 million Claude interactions in the largest AI distillation attack on record.

Midjourney has asked a federal judge to compel Disney, Universal, and Warner Bros. to disclose their own AI training practices - using the studios' conduct as a fair-use and unclean-hands defense in a high-stakes copyright case.