Recent Articles - Page 8

Latest News

VideoVerse's $250M Exit Collapsed Into Fraud Suits

VideoVerse's $250M Exit Collapsed Into Fraud Suits

Minute Media's purchase of AI video startup VideoVerse fell apart after the deal closed, and three separate Delaware lawsuits now accuse founder Vinayak Shrivastav of forging signatures to extract tens of millions.

Model Steering, Angry Buyers, and Blind Judges

Model Steering, Angry Buyers, and Blind Judges

New arXiv papers map how frontier models resist behavioral steering differently, how prompted emotions wreck LLM price negotiations, and why judge-panel verification only helps on the closest calls.

View All News →

Guides

View All →

Reviews

View All →

Leaderboards

View All →

Models

View All →
GPT-5 mini

GPT-5 mini

OpenAI's original cost-efficient GPT-5 variant pairs a 400K context window with $0.25/$2.00 per million token pricing, still doing quiet duty as a cheap backbone for research agents a year after launch.

NVIDIA Nemotron 3.5 Lightning 30B-A3B

NVIDIA Nemotron 3.5 Lightning 30B-A3B

NVIDIA's 30B MoE model with 3B active parameters, distilled from Nemotron 3 Ultra, hits 86% PinchBench accuracy at up to 4x the output speed of comparable open models.

Qwen3.8-Max

Qwen3.8-Max

Alibaba's 2.4 trillion parameter flagship ships with real pricing and a published benchmark table, but the open-weight release it promised for this week still hasn't shown up.

Recent

DeepSeek-R1

DeepSeek-R1

DeepSeek-R1 is the 671B-parameter open-weight reasoning model that matched OpenAI o1 on math and coding benchmarks and triggered a $589 billion single-day drop in Nvidia's market cap in January 2025.

Anthropic's $1.5B Book Piracy Settlement Wins Approval

Anthropic's $1.5B Book Piracy Settlement Wins Approval

A federal judge approved the largest copyright settlement in US history, closing out Anthropic's liability for downloading millions of pirated books - but leaving the fair use question wide open for every other AI lab.

Google's Frozen v2 Chip Bakes Gemini Into Silicon

Google's Frozen v2 Chip Bakes Gemini Into Silicon

Google is reportedly building a chip line separate from its TPUs that hardwires parts of Gemini directly into silicon, promising up to 10x efficiency as a capacity crunch forces Cloud to turn away customers.

Two World Models, One Multi-Agent Review Problem

Two World Models, One Multi-Agent Review Problem

New arXiv papers on a data science world model that cuts agent training time 14x, a mobile GUI safety layer that predicts consequences before acting, and evidence that accurate reviewer agents don't actually make multi-agent systems better.

Luma Ray3.2

Luma Ray3.2

Luma Ray3.2 is Luma AI's current flagship video model - native 16-bit HDR, 16-keyframe control, and the company's first full developer API, but still no native audio.

Pika 2.5

Pika 2.5

Pika Labs' flagship video model trades cinematic Elo rankings for the deepest creative-effects toolkit in AI video, plus a pivot into real-time agent video with PikaStream.

Haiper 2.x

Haiper 2.x

Haiper 2.x is the cheapest per-second AI video API on the market at $0.033/sec, now run by NetMind.AI after Haiper's consumer app shut down and its founders joined Microsoft.