Browse guides

AI Models articles

38
Claude AI

Claude Pricing 2026: Every Model, Every Tier, Full Breakdown

Every Claude model and pricing tier in one guide. API pricing for Fable 5 ($10/$50), Opus 5 ($5/$25), Sonnet 5 ($3/$15 intro $2/$10), and Haiku 4.5 ($1/$5) per MTok. Subscription plans from Free to Enterprise. Prompt caching, batch API, fast mode, and cross-provider comparisons — all verified July 25, 2026.

Grok 4.5

Grok 4.5 Cracks OpenRouter's Top 10 Closed Models: What the Token-Volume Milestone Actually Means

Polymarket reports Grok 4.5 has overtaken GPT-5.6 Sol and Fable 5 in OpenRouter token volume, cracking the top 10 closed models. The milestone measures usage, not quality — but it lines up with a genuinely aggressive value proposition: $2/$6 per 1M tokens, roughly 2x token efficiency, 80 TPS, a SWE Marathon lead, and Cursor-co-trained coding performance. Here is what the signal means, what xAI’s benchmarks show, and what they do not.

FLUX 3

FLUX 3 Released: Black Forest Labs Turns Its Image Model Into a Multimodal Video, Audio and Robotics Engine

Black Forest Labs launched FLUX 3 on July 23, 2026, in Early Access. Unlike FLUX 1 and FLUX 2, which generate images, FLUX 3 is a single multimodal flow model trained jointly on images, video, and audio — and even robot actions. It generates video with native audio up to 20 seconds, and an early version already powers FLUX-mimic, a video-action robotics model deployed at Audi. Here is what FLUX 3 does, what is still preliminary, and what the rollout plan covers.

Sakana AI

Fugu-Ultra v1.1 Released: Sakana AI's Orchestration Model Gets a Frontier Refresh

Sakana AI shipped Fugu-Ultra v1.1 on July 24, 2026, refreshing its multi-agent orchestration model with newer frontier workers. Sakana claims up to 7.9 benchmark points over v1.0, with the biggest gains on ProgramBench and Terminal Bench 2.1, and keeps pricing identical to v1.0. Here is what changed, what is still vendor-reported, and how to evaluate the upgrade.

Claude AI

Claude Opus 5: Benchmarks, Pricing, and Full Guide (July 2026)

Claude Opus 5 launched July 24, 2026. It delivers near-Fable 5 intelligence at Opus pricing ($5/$25 per MTok), with a 1M-token context window, 128K max output, and state-of-the-art results on Frontier-Bench, ARC-AGI 3, and OSWorld 2.0. This guide covers benchmarks, pricing, safety, access, and how Opus 5 compares to Fable 5, Sonnet 5, and the competition.

Gemini 4

Gemini 4 Training Has Begun: What Google Confirmed—and What It Did Not

Google has started what it calls its most ambitious pre-training run yet for Gemini 4, but has not announced a release date, model lineup, benchmarks, pricing, context window, or public preview.

Gemini AI

Gemini 3.6 Flash Launch: Price, Benchmarks, API & Flash-Lite

Google’s July 21 Gemini update adds a more token-efficient 3.6 Flash, a 350-token-per-second 3.5 Flash-Lite, and the restricted Gemini 3.5 Flash Cyber model for CodeMender.

Qwen 3.8

Qwen 3.8: Preview Access, Specs, Pricing & Benchmarks

A source-checked guide to Qwen3.8-Max-Preview covering release status, 2.4T scale, context and output limits, reasoning controls, Qwen Cloud pricing, early coding evidence, and the promised open-weight release.

Kimi K3

Kimi K3 Sold Out: Why Moonshot Paused New Subscriptions

Moonshot AI paused new Kimi subscriptions after Kimi K3 demand pushed GPU capacity close to the limit. Here is what sold out means, what existing users keep, and how Kimi K3 compares with Claude Opus 4.8 and GPT-5.6 Sol.

Reasoning Models

What is a Reasoning Model?

A practical guide to reasoning models: what separates them from standard language models, how inference-time deliberation works, the tasks that benefit most, common failure modes, and a seven-step checklist for prompting and verifying results.