Models

Hermes 4.3

Hermes 4.3

Nous Research's 36B open-weight model matches Hermes 4 70B on most benchmarks, tops RefusalBench on alignment, and is the first production model trained entirely on the Solana-secured Psyche network.

Gemini 2.5 Flash

Gemini 2.5 Flash

Google's hybrid reasoning workhorse pairs a 1M-token context window with $0.30/$2.50 per million token pricing and a toggleable 0-24,576 token thinking budget, now heading toward an October 2026 shutdown.

Seedream 5.0 Pro

Seedream 5.0 Pro

ByteDance's July 2026 flagship image generation model with native 2K output, editable layer separation, multilingual text in 14 languages, and precision region editing.

Grok 4.5

Grok 4.5

Grok 4.5 is xAI's 1.5-trillion-parameter V9 MoE model, publicly launched July 8 at $2/M input - cheap, fast, and token-efficient, though neutral harness benchmarks put it well behind Fable 5 and Opus 4.8 on coding.

GPT-Live-1

GPT-Live-1

OpenAI's full-duplex voice model that listens and speaks simultaneously, replacing Advanced Voice Mode in ChatGPT with three reasoning tiers backed by GPT-5.5.

Grok 4.1 Fast

Grok 4.1 Fast

Grok 4.1 Fast is xAI's agent-optimized model with a 2M-token context window, #1 ranking on tau-bench Telecom, and one of the lowest input prices among frontier-adjacent APIs at $0.20/M tokens.

GPT-5.4 mini

GPT-5.4 mini

OpenAI's mid-range model in the GPT-5.4 family delivers near-flagship coding and agentic performance at $0.75/M input tokens with a 400K context window.

LongCat-2.0

LongCat-2.0

Meituan's 1.6T-parameter open-source MoE coding model, trained end-to-end on 50,000 domestic Chinese ASICs, with native 1M token context and a 59.5 SWE-bench Pro score.

Gemini Omni Flash

Gemini Omni Flash

Google DeepMind's multimodal video generation model that creates 10-second clips with native audio from text, images, or video inputs - and lets you refine results through conversation.

Holo3-35B-A3B

Holo3-35B-A3B

H Company's open-weight sparse MoE vision-language model purpose-built for desktop computer use, scoring 82.6% on OSWorld-Verified with only 3B active parameters.

Claude Sonnet 5

Claude Sonnet 5

Anthropic's latest Sonnet-class model brings near-Opus coding performance to mid-tier pricing, with major agentic search and computer use gains over Sonnet 4.6.

GPT-5.6

GPT-5.6

OpenAI's GPT-5.6 family - Sol, Terra, and Luna - sets a new Terminal-Bench 2.1 record at 91.9% with subagent Ultra mode, but remains locked to ~20 government-vetted partners as of launch.