Recent Articles - Page 85

Latest News

VideoVerse's $250M Exit Collapsed Into Fraud Suits

VideoVerse's $250M Exit Collapsed Into Fraud Suits

Minute Media's purchase of AI video startup VideoVerse fell apart after the deal closed, and three separate Delaware lawsuits now accuse founder Vinayak Shrivastav of forging signatures to extract tens of millions.

Model Steering, Angry Buyers, and Blind Judges

Model Steering, Angry Buyers, and Blind Judges

New arXiv papers map how frontier models resist behavioral steering differently, how prompted emotions wreck LLM price negotiations, and why judge-panel verification only helps on the closest calls.

View All News →

Guides

View All →

Reviews

View All →

Leaderboards

View All →

Models

View All →
GPT-5 mini

GPT-5 mini

OpenAI's original cost-efficient GPT-5 variant pairs a 400K context window with $0.25/$2.00 per million token pricing, still doing quiet duty as a cheap backbone for research agents a year after launch.

NVIDIA Nemotron 3.5 Lightning 30B-A3B

NVIDIA Nemotron 3.5 Lightning 30B-A3B

NVIDIA's 30B MoE model with 3B active parameters, distilled from Nemotron 3 Ultra, hits 86% PinchBench accuracy at up to 4x the output speed of comparable open models.

Qwen3.8-Max

Qwen3.8-Max

Alibaba's 2.4 trillion parameter flagship ships with real pricing and a published benchmark table, but the open-weight release it promised for this week still hasn't shown up.

Recent

Stanford's AI Index 2026 - US Edge Over China Is Gone

Stanford's AI Index 2026 - US Edge Over China Is Gone

Stanford HAI's 2026 AI Index finds the US-China model gap has effectively closed, GenAI has hit 53% global adoption faster than any prior technology, and young software developers are the first casualties of the labor shift.

The AI Layoff Trap - Game Theory Says Everyone Loses

The AI Layoff Trap - Game Theory Says Everyone Loses

A UPenn-BU paper models AI-driven layoffs as a Prisoner's Dilemma: each firm wins by automating, but when everyone does it, collapsing demand makes every firm worse off. Their proposed fix is a Pigouvian tax on automated tasks.

Claude Code Silently Burns 40% More Tokens Since v2.1.100

Claude Code Silently Burns 40% More Tokens Since v2.1.100

A developer used an HTTP proxy to capture full API requests across four Claude Code versions and found that v2.1.100 adds roughly 20,000 invisible server-side tokens to every request - inflating billing by 40% with no user visibility.

llama.cpp Lands Three Audio Models in 48 Hours

llama.cpp Lands Three Audio Models in 48 Hours

Three separate PRs merged into llama.cpp between April 11-13 add MERaLiON-2, Gemma 4's Conformer encoder, and Qwen3-Omni/ASR - making local voice AI inference practical on consumer hardware for the first time.