Model names refreshed to the September 2026 lineup. The benchmark figures deliberately keep the model they were measured on, because those tests have not been re-run on the newer models; they are now dated so the age is visible.
What changed, and when
Every update to every solution and model entry. Newest first.
Updates
Cost rows updated to Claude Sonnet 5 ($2/$10) and GPT-5.6 Terra ($2/$12), replacing Sonnet 4.6 and the deprecated GPT-4o. The Sonnet per-character estimate only fell to ~$3.50 rather than by the full price cut, because Claude 5's tokenizer produces roughly 30% more tokens for the same text.
Verified against claude.com/pricing. Corrected the Claude Team price — it is $20/seat annually, not $25 — and added the Team Premium tier at $100/seat. Updated the model-under-the-hood row from Sonnet 4.6 / Opus 4.7 to Sonnet 5 / Opus 5, noting that Opus requires a paid plan. Jasper, Copy.ai and Writer pricing was not re-verified this round.
Model names in the inference-cost row updated to Claude Sonnet 5 and GPT-5.6 Terra. GPT-4o was named here and is deprecated, superseded by GPT-5.5. The cost range is unchanged. Added referenced-model links for models this page already names: Claude Sonnet 5, GPT-5.6 Terra.
Referenced-model links updated to the current lineup: Gemini 3.1 Flash-Lite to Gemini 3.5 Flash-Lite. Page content not otherwise re-verified.
Referenced-model links updated to the current lineup: Gemini 3.1 Flash-Lite to Gemini 3.5 Flash-Lite. Page content not otherwise re-verified.
Referenced-model links updated to the current lineup: Gemini 3.1 Flash-Lite to Gemini 3.5 Flash-Lite. Page content not otherwise re-verified.
Two corrections. Claude's context window is set by the model, not the plan: any paid plan gets 1M when chatting with Opus 5, Sonnet 5 or Fable 5.1, and older models give 500K or 200K. And Microsoft names no model at all, routing automatically per prompt. Flagship models updated to Claude Opus 5 and Sonnet 5.
Corrected the model prerequisite. It described Claude as '200k default, 1M on the higher tiers', which is backwards for the current generation: every Claude 5 model is 1M with no smaller variant to fall back to. The deprecated GPT-4o reference was replaced.
Referenced-model links updated to the current lineup: Gemini 3.1 Flash-Lite to Gemini 3.5 Flash-Lite. Page content not otherwise re-verified.
Model and plan rows re-verified against each vendor's own docs. Cursor does not offer GPT-6 Astra; GitHub Copilot has added a Max tier at $100/month; Claude Code defaults to Opus 5 on Max and above and Sonnet 5 on Pro. Claude Pro corrected to $17/month billed annually.
Recommended default models updated from Claude Sonnet 4.6 and GPT-4o to Claude Sonnet 5 and GPT-5.6 Terra, in both the prose and the cost row. GPT-4o is deprecated and two generations behind, so the old recommendation was steering readers at a model OpenAI has superseded. Added referenced-model links for models this page already names: Claude Sonnet 5, GPT-5.6 Terra.
Referenced-model links updated to the current lineup: Gemini 3.1 Flash-Lite to Gemini 3.5 Flash-Lite. Page content not otherwise re-verified.
Updated the Claude Vision reference from Sonnet 4.6 / Opus 4.7 to Sonnet 5 / Opus 5. Repointed the referenced-models links to Mistral OCR 4.1 and the current Claude and GPT entries.
Referenced-model links updated to the current lineup: Gemini 3.1 Flash-Lite to Gemini 3.5 Flash-Lite. Added referenced-model links for models this page already names: GPT-4.1, OpenAI text-embedding-3-large. Page content not otherwise re-verified.
Referenced-model links updated to the current lineup: Gemini 3.1 Flash-Lite to Gemini 3.5 Flash-Lite. Added referenced-model links for models this page already names: GPT-4.1. Page content not otherwise re-verified.
Added referenced-model links for models this page already names: BGE-M3, Cohere Embed v4, OpenAI text-embedding-3-large, OpenAI text-embedding-3-small. Page content not otherwise re-verified.
Added referenced-model links for models this page already names: OpenAI text-embedding-3-large, OpenAI text-embedding-3-small. Page content not otherwise re-verified.
Referenced-model links updated to the current lineup: Gemini 3.1 Flash-Lite to Gemini 3.5 Flash-Lite. Page content not otherwise re-verified.
Flagship-model cost row updated to Claude Sonnet 5 / GPT-5.6 Terra / Gemini 3.1, and the referenced models repointed to Mistral OCR 4.1 and the current Claude and GPT entries. The per-page cost range is unchanged.
Referenced-model links updated to the current lineup: Gemini 3.1 Flash-Lite to Gemini 3.5 Flash-Lite. Page content not otherwise re-verified.
Referenced-model links updated to the current lineup: Claude 4.6 Sonnet to Claude Sonnet 5; GPT-4o to GPT-5.6 Terra. Added referenced-model links for models this page already names: OpenAI text-embedding-3-small. Page content not otherwise re-verified.
Updated the per-image vision cost row to Claude Sonnet 5 / GPT-5.6 Terra / Gemini 3.1. The cost range is unchanged. Referenced-model links updated to the current lineup: GPT-4o to GPT-5.6 Terra; Claude 4.7 Opus to Claude Opus 5. Added referenced-model links for models this page already names: Claude Sonnet 5, GPT-5.5.
Referenced-model links updated to the current lineup: Gemini 3.1 Flash-Lite to Gemini 3.5 Flash-Lite. Page content not otherwise re-verified.
Pricing, context windows and model names refreshed to the September 2026 lineup. The quality rankings still come from the Q1 2026 blind evaluation of the versions current at the May 2026 snapshot; that has not been re-run, and the page now says so up front.
Referenced-model links updated to the current lineup: Gemini 3.1 Flash-Lite to Gemini 3.5 Flash-Lite. Page content not otherwise re-verified.
Added referenced-model links for models this page already names: Gemini 2.5 Flash. Page content not otherwise re-verified.
Referenced-model links updated to the current lineup: Gemini 3.1 Flash-Lite to Gemini 3.5 Flash-Lite. Page content not otherwise re-verified.
Referenced-model links updated to the current lineup: Gemini 3.1 Flash-Lite to Gemini 3.5 Flash-Lite. Page content not otherwise re-verified.
Frontier and open-weight model references refreshed to the September 2026 lineup. The open-weight row is now tiered by hardware, because naming DeepSeek V4 Pro flat in the 'Local' column implied a 1.6T model runs on the workstation budgeted two rows below.
Referenced-model links updated to the current lineup: Claude 4.7 Opus to Claude Opus 5; GPT-4o to GPT-5.6 Terra. Page content not otherwise re-verified.
Referenced-model links updated to the current lineup: Claude 4.6 Sonnet to Claude Sonnet 5. Page content not otherwise re-verified.
Updated the proprietary-flagship price stat to the September 2026 lineup (GPT-6 Astra, Claude Opus 5, Gemini 3.1 Pro). The range moved from $3–$15 to $2–$10 per million input tokens. Added referenced-model links for models this page already names: Claude Haiku 4.5, Claude Opus 5, Gemini 3.5 Flash-Lite, GPT-5.6 Luna, GPT-6 Astra.
The model-choice step now leads with Qwen3.8 27B, which is Apache 2.0 and fits the single-GPU pilot tier, instead of Llama 3.3 70B, which pushed readers toward far more expensive hardware than the job needs. The measured latency figures stay with the model they were measured on.
Referenced-model links updated to the current lineup: Gemini 3.1 Flash-Lite to Gemini 3.5 Flash-Lite. Page content not otherwise re-verified.
Cost estimate moved from Claude Sonnet 4.6 to Sonnet 5 pricing. The range tightened only slightly (~$0.01-$0.035 from ~$0.01-$0.04) rather than falling by the full third the price cut implies, because Claude 5's tokenizer produces roughly 30% more tokens for the same text. Added referenced-model links for models this page already names: Claude Sonnet 5.
Referenced-model links updated to the current lineup: Gemini 3.1 Flash-Lite to Gemini 3.5 Flash-Lite. Page content not otherwise re-verified.
Referenced-model links updated to the current lineup: Claude 4.7 Opus to Claude Opus 5. Page content not otherwise re-verified.
Corrected the model-API prerequisite from 'Claude (200k/1M), GPT (128k-272k)' to the current windows: Claude 5 is 1M on every model, GPT-5.6 and GPT-6 reach ~1.05M. last_verified deliberately not bumped: only this line was re-checked, not the whole page. Referenced-model links updated to the current lineup: Gemini 3.1 Flash-Lite to Gemini 3.5 Flash-Lite.
Per-summary cost recalculated at Claude Sonnet 5 pricing: about $0.03, down from $0.05. Claude's context figure corrected too, since any paid plan chatting with Opus 5, Sonnet 5 or Fable 5.1 gets the full 1M.
New section on what actually sets your context window: the model you pick, not the plan you pay for. Any paid Claude plan gets 1M on Opus 5, Sonnet 5 or Fable 5.1, while older models give 500K or 200K. Pricing refreshed to the September 2026 lineup and both worked cost examples recalculated.
Referenced-model links updated to the current lineup: Gemini 3.1 Flash-Lite to Gemini 3.5 Flash-Lite. Added referenced-model links for models this page already names: GPT-4.1. Page content not otherwise re-verified.
Added referenced-model links for models this page already names: BGE-M3. Page content not otherwise re-verified.
Referenced-model links updated to the current lineup: OpenAI TTS to GPT Transcribe. Page content not otherwise re-verified.
Updated the agent-capability example from Claude Sonnet 4.5+/Opus 4.7 to Claude Sonnet 5 / Opus 5. Added referenced-model links for models this page already names: Claude Sonnet 5.
Referenced-model links updated to the current lineup: GPT-4o to GPT-5.6 Terra; Claude 4.7 Opus to Claude Opus 5. Added referenced-model links for models this page already names: GPT-5.5. Page content not otherwise re-verified.
Added referenced-model links for models this page already names: Distil-Whisper. Page content not otherwise re-verified.
Verified current. No bge-m4 exists; BGE-M3's successors are task-specific siblings rather than a new generation.
Verified current. No bge-reranker-v3 exists. The newer v2.5-gemma2-lightweight is a heavier variant under the restrictive Gemma licence, not a general replacement.
Deprecated — superseded by Claude Sonnet 5 (released 2026-06-30), documented by Anthropic as a drop-in replacement. Sonnet 4.6 moves to the legacy table; retirement not sooner than 2027-02-17.
Deprecated — superseded by Claude Fable 5.1 (released 2026-09-01). Fable 5 moves to Anthropic's legacy table; still available, retirement not sooner than 2027-06-09.
Initial entry. Released 2026-09-01 as the successor to Claude Fable 5, at the same $10/$50 per MTok with cache reads cut to $0.25 per MTok. Adds adaptive thinking always on; drops support for forced tool use.
Verified current against Anthropic's models overview. Still a column in the current-lineup table; no Haiku 5 exists. Retirement not sooner than 2026-10-15.
Deprecated — superseded by Claude Mythos 5.1 (released 2026-09-01). Still in limited availability; not listed on Anthropic's deprecation table.
Initial entry. Released 2026-09-01 alongside Claude Fable 5.1, succeeding Claude Mythos 5. Remains invitation-only via Project Glasswing.
Deprecated — superseded by Claude Opus 5 (released 2026-07-24). Opus 4.8 moves to Anthropic's legacy table; still available on all platforms, retirement not sooner than 2027-05-28.
Initial entry. Released 2026-07-24 as the Opus-tier flagship, superseding Claude Opus 4.8. 1M-token context (default and maximum), $5/$25 per MTok, adaptive thinking on by default with a low/medium/high/xhigh/max effort ladder.
Initial entry. Released 2026-06-30 as a drop-in replacement for Claude Sonnet 4.6. The $2/$10 per MTok launch pricing became the standard price — the increase to $3/$15 scheduled for 2026-09-01 was cancelled.
Verified current against Mistral's models overview — the only model in the Code Models section, and not on the deprecation table.
Verified current against Cohere's models page. embed-v4.0 is still the newest Embed generation; everything below is v3.0/v2.0. No v5 exists.
Verified current against Cohere's models page and the Rerank v4.0 changelog. rerank-v4.0-pro is top of the rerank list; rerank-v4.0-fast is a sibling tier, not a successor.
Verified: the weights remain hosted, ungated and heavily used (7.2M downloads last month) despite Coqui the company being defunct — this entry's existing 'community maintenance only' weakness is accurate. Maintenance has moved to idiap/coqui-ai-TTS.
Corrected on three counts. Release date 2026-05-06 was a Hugging Face push timestamp, not a release — the official DeepSeek-V4-Flash-0731 shipped 2026-07-31 per DeepSeek's API changelog, superseding the preview this entry described. Licence corrected from 'DeepSeek License' to MIT, and the false 'restrictions on commercial use cases' weakness removed. Added parameter size and 1M context window.
Corrected on three counts. Release date 2026-05-06 was a Hugging Face push timestamp, not a release — the GA DeepSeek-V4-Pro-0813 shipped 2026-08-13 per DeepSeek's API changelog, superseding the preview this entry described. Licence corrected from 'DeepSeek License' to MIT, and the false 'restrictions on commercial use cases' weakness removed. Added parameter size and 1M context window.
Deprecated — Mistral's deprecation table shows devstral-2512 deprecated 2026-05-22 and retired 2026-07-31, with Mistral Medium 3.5 as the stated alternative. The entry had been showing a retired model as current.
Verified current against the full distil-whisper org (12 models). The distil-large-v3.5 family is newest; no v4 exists. Worth knowing: distil-large-v3 still has roughly 90x the downloads of v3.5.
Dropped Turbo v2.5 from purpose and strengths — ElevenLabs' models page now marks Eleven Turbo v2.5 and Turbo v2 as deprecated, 'outclassed by Flash models', with eleven_flash_v2_5 as the replacement. Eleven v3 itself verified current; no v4 exists.
Verified current against BFL's FLUX.2 launch post. No newer [pro] version exists; FLUX.2 [max] is a sibling tier, not a successor.
Verified current against Black Forest Labs' docs, which state FLUX.2 remains fully supported for production. FLUX 3 exists but its Dev open-weight backbone has not shipped.
Verified current against Google's models and deprecations pages. Still GA and priced; no shutdown announced. Generations behind the Flash line but not superseded within its own 2.5 family.
Verified current against Google's models and deprecations pages. Still GA with no shutdown announced — and now the only GA (non-preview) Pro model, since Gemini 3.5 Pro was shelved.
Deprecated by Google — the deprecations page gives a 2027-05-07 shutdown and names Gemini 3.5 Flash-Lite (2026-07-21) as the replacement.
Release date corrected to 2026-03-11 from Google's deprecations page. The previous 2026-05-07 was the models page's global 'last updated' stamp. Verified still current with no shutdown announced.
Verified current. Still the top Pro model on Google's docs — Gemini 3.5 Pro was announced at I/O, missed three times and reportedly shelved to pretrain Gemini 4. No Pro model newer than this exists.
Deprecated — Google's models page now describes it as the legacy Flash model. Superseded by Gemini 3.8 Flash (2026-09-02), with 3.6 and 3.7 Flash released in between. No shutdown date announced.
Initial entry. Released 2026-07-21; named on Google's deprecations page as the replacement for Gemini 3.1 Flash-Lite (shutdown 2027-05-07). Context window left unset — Google publishes none.
Initial entry. Released 2026-09-02 per Google's Gemini API changelog as the head of the Flash line, superseding Gemini 3.7 Flash (2026-08-13) and 3.6 Flash (2026-07-21). Dates taken from the changelog and deprecations pages, not the models page (whose per-model stamps are a global page timestamp).
Release date corrected to 2025-12-17 from Google's deprecations page; the previous 2026-05-07 was the models page's global 'last updated' stamp. Note Google names gemini-3.6-flash as the official replacement, though the catalog chains it through Gemini 3.5 Flash.
Verified current against Google's Gemma pages. Gemma 4 (April 2026) is the newest generation and 31B the top core size. Gemma is absent from the Gemini API models page by design — that is a coverage gap, not a deprecation.
Initial entry — first Z.ai model in the catalog. Released 2026-08-26, eight days after the flagship GLM-5.3 (2026-08-18). The Flash variant is catalogued rather than the flagship because it is plain MIT, while GLM-5.3 carries a custom licence requiring a Z.ai security review for MaaS operators above $10B revenue. Pricing recorded at list ($0.15/$0.50); a 50% promotional rate expires 2026-09-09. Release date taken from Z.ai's release notes, NOT from the arXiv paper linked on the model card — that paper (2026-02-17) is the shared GLM-5 base-model report, not this model's release.
Verified current against OpenAI's models overview — still listed under speech.
Deprecated by OpenAI — announced 2026-08-26, shutdown 2027-02-26, replaced by gpt-transcribe and gpt-live-transcribe.
Pricing corrected to $0.20 input / $1.20 output per MTok — OpenAI cut Luna's rates about 80% on 2026-07-30; the entry carried the July launch prices.
Pricing corrected to $4 input / $20 output per MTok — OpenAI cut Sol's rates on 2026-08-21 (promotional at least through 2026-11-21); the entry carried the July launch prices. Repositioned below GPT-6 Astra (released 2026-09-03).
Pricing corrected to $2 input / $12 output per MTok — OpenAI cut Terra's rates roughly 20% on 2026-07-30; the entry carried the July launch prices.
Initial entry. Released 2026-09-03 per OpenAI's API changelog as 'our most capable model, built for the hardest end-to-end work'. Specs and pricing verified against the gpt-6-astra model docs.
Verified current against OpenAI's models overview — still listed, snapshot gpt-image-2-2026-04-21.
Initial entry. Released 2026-07-28. Named alongside gpt-transcribe on OpenAI's deprecations page as a replacement for gpt-4o-transcribe (shutdown 2027-02-26). Context window and max output are not published by OpenAI.
Initial entry. Released 2026-07-28. Named on OpenAI's deprecations page as the replacement for gpt-4o-transcribe (shutdown 2027-02-26). Context window and max output are not published by OpenAI, so both are left unset.
Initial entry — first SpaceXAI model in the catalog. Released 2026-08-12. Maker recorded as SpaceXAI: that is the string on the copyright notice, the nav on every announcement page, the model docs' own prose, and the Arena leaderboard's provider column. The company was formerly xAI, and the API and SDK identifiers are still xai_sdk / XAI_API_KEY / api.x.ai. Reasoning controls run low / medium / high / xhigh with no way to disable.
Initial entry. Released 2026-08-07. SpaceXAI's announcement claims #2 worldwide on the Arena text-to-image and image-edit leaderboards as of the launch date; the live text-to-image leaderboard on 2026-09-04 shows it at #3, displaced by Microsoft's mai-image-2.6, with gpt-image-2 still #1. The entry records the current position rather than the launch claim. Current Image Edit standing was not re-checked.
Initial entry — first Moonshot AI model in the catalog. Weights released 2026-07-27; the API launched 2026-07-17. The licence was read in full: contrary to widely repeated secondary reports, it contains NO revenue-sharing clause and no percentage. What it has is a negotiation gate — MaaS operators above $20M revenue over any rolling 12 months must agree separate terms first — plus an attribution requirement above 100M MAU or $20M monthly revenue. Internal use and access through Moonshot's own products are exempt.
Verified current against Kuaishou's own Kling 3.0 quickstart guide. Nothing newer than the 3.0 series is named; all 'Kling 4.0' material is forward-looking speculation.
Release date corrected to 2025-04-05, Meta's actual Llama 4 launch. The previous 2025-05-22 was the Hugging Face repo push timestamp. Still the newest Llama — there is no Llama 5; Meta's open-weight work moved to the Muse family (see Muse Glimmer 30B).
Release date corrected to 2025-04-05, Meta's actual Llama 4 launch. The previous 2025-05-22 was the Hugging Face repo push timestamp. Still the newest Llama — there is no Llama 5; Meta's open-weight work moved to the Muse family (see Muse Glimmer 30B).
Updated from Luma Ray3 to Ray3.2 (released 2026-06-09), confirmed as the current model on lumalabs.ai/llm-info. Ray3.14 shipped 2026-01-26 in between. Note for future runs: docs.lumalabs.ai/docs/video-generation is a generation behind (it documents only the deprecated ray-2 line) — use lumalabs.ai/llm-info as the verification source instead. No Ray4 exists.
Verified current against Mistral's models overview and model card; release date matches exactly.
Verified current against Mistral's models overview and model card; release date matches exactly.
Release date corrected to 2026-04-28 from Mistral's own model card and changelog; the previous 2026-04-01 was inferred from the 26.04 version stamp. Verified current — Mistral now names it the alternative for both the retired Devstral 2 and the retired Magistral line, with reasoning exposed through the reasoning_effort parameter rather than a separate model.
Verified current against Mistral's models overview and model card; release date matches exactly. Now described as a hybrid model unifying instruct, reasoning and coding.
Initial entry. Released 2026-08-10 under the new meta-models Hugging Face org. Meta's open-weight work has moved off the Llama name — there is no Llama 5, and the closed-weight Muse Spark (2026-04-08) replaced Llama as Meta's flagship. Llama 3.1 8B is kept in the catalog because Llama 4's smallest model is 109B, leaving Muse Glimmer as the only current Meta option for a single consumer GPU.
Verified current — no Parakeet v4 exists. Note nvidia/parakeet-unified-en-0.6b (2026) now beats this model both offline and dramatically in streaming, so the 'top-of-leaderboard English accuracy' claim needs a re-check against the current Open ASR leaderboard.
Superseded by OCR 4.1. Release date corrected to 2025-12-18 to match Mistral's model card. Still available and not on Mistral's deprecation table.
Initial entry. Released 2026-07-16 and generally available 2026-08-31 per Mistral's changelog; mistral-ocr-latest and mistral-ocr-4 now point at it. OCR 4.0 shipped nine weeks earlier and is skipped as an interim version.
Verified current against the model card — 14B, 16K context, MIT, no deprecation notice. Phi-5 does not exist.
Pika 2.5 confirmed as the current model — Pika's own developer docs list pika/pika-2.5/text-to-video and image-to-video with published pricing ($0.04/sec at 720p, $0.09/sec at 1080p, 5s text-to-video cap). Entry held at 2.2 because Pika publishes no release date for 2.5 anywhere, and this catalog does not guess release dates. Bump the title once a dated source appears.
Verified actively maintained: OHF-Voice/piper1-gpl shipped v1.8.0 on 2026-09-04, with v1.5.0 through v1.7.0 across July and August 2026. The Open Home Foundation's README notes it is seeking maintainers.
Superseded by Qwen3.8 27B (August 2026). Note the filename is a legacy artifact of the corrected 2026-06-08 entry and no longer matches the title.
Initial entry. Qwen3.8 released August 2026, superseding Qwen3.6 (April 2026) with Qwen3.5 (February 2026) in between. Alibaba's model card gives only a month, so the release date is recorded as the first of the month rather than a guessed day. Benchmarks and licence taken from the model card, not from Hugging Face push timestamps.
Verified current and not archived; last release v0.3.0 (2022-09-20). No successor exists — the README's GFPGAN/BasicSR links are complementary projects, not replacements. Canonical repo remains under the author's personal account; TencentARC/Real-ESRGAN returns 404.
Verified current against Runway's model docs and changelog, which dates Gen-4.5 to 2025-12-11 — matching this entry exactly. No Gen-5 exists. Aggregator claims of a March 2026 launch contradict Runway's own changelog and were rejected.
Verified current against Stability's news page, which dates Stable Audio 3.0 to 2026-05-20 — matching this entry exactly.
Verified current. A full Hugging Face org enumeration by creation date shows no Stable Diffusion 4 repo, and Stability's own 2026 news page never mentions one — the widely circulated 'SD4, April 2026' claim is fabricated.
Verified current — no newer open-weights Stability upscaler exists in the full org listing.
Verified current. No StarCoder3 exists. BigCode still ships evaluation datasets but has published no new model since 2024, so dormancy here is not staleness.
Verified current against suno.com and Suno's release notes, which confirm v5.5 shipped 2026-03-26 — matching this entry exactly. No v6 exists. Watch item: Suno has said current models will be deprecated when its Warner/BMG-licensed models ship later in 2026.
Verified current as the canonical single-image restoration implementation. Repo is not archived; last real commit 2022-12-04. The authors' VRT/RVRT work is video restoration — adjacent, explicitly not a replacement.
Corrected on four counts. Renamed from 'Topaz Gigapixel AI' to 'Topaz Gigapixel' — Topaz dropped the 'AI' suffix. Release date moved from 2018-06-01 to 2025-09-16, the current product's date per the product page's schema markup; the 2018 date predated the current product entirely. Product URL updated from /gigapixel-ai to /gigapixel. Pricing corrected from a one-time licence to subscription-only. Version 1.3.3 shipped 2026-07-27.
Verified current against Udio's own help-centre changelog, which names Allegro v1.5 (2025-03-18) as the newest model — matching this entry exactly. Third-party claims of a 'Udio v4' or 'v3.5' have no official backing anywhere and were rejected again this run.
Release date corrected to 2025-10-15 from Google's deprecations page. The previous 2026-05-07 was the models page's global 'last updated' stamp, not a release date. Verified still current — no Veo 4 exists.
Release date corrected to 2026-03-31 from Google's deprecations page. The previous 2026-05-07 was the models page's global 'last updated' stamp, not a release date. Verified still current.
Verified actively maintained — real feature commits through 2026-08-09, despite no tagged release since v0.3.50 (2024-04-22). The release gap is not staleness. No successor project exists.
Verified current against Voyage's embeddings docs. voyage-4-large is still the best general-purpose retrieval model. Note voyage-context-4 is a different family (contextualised chunk embeddings), not a newer version of this one.
Link fix: OpenAI's image guide moved — /api/docs/guides/images now 404s. Repointed the provider and official_link to /api/docs/guides/image-generation.
Link fix: promoted the models overview (elevenlabs.io/docs/overview/models) to the primary docs link. The API reference was the only docs link, and it does not enumerate models — which made it a poor verification source. Kept as a secondary link.
Verified current after the GPT-5.6 launch — GPT-5.6 is an additional frontier tier, not a replacement. OpenAI's deprecations page does not list GPT-5.5, and still names it as the recommended successor for o3, GPT-5, and the chat-latest snapshots. Release date corrected 2026-04-23 → 2026-04-24 per the API changelog.
Initial entry. GPT-5.6 family released 2026-07-09 per OpenAI's API changelog; specs and pricing verified against the gpt-5.6-luna model docs.
Initial entry. GPT-5.6 family released 2026-07-09 per OpenAI's API changelog; specs and pricing verified against the gpt-5.6-sol model docs. API alias is `gpt-5.6`.
Initial entry. GPT-5.6 family released 2026-07-09 per OpenAI's API changelog; specs and pricing verified against the gpt-5.6-terra model docs.
Deprecated — superseded by Mistral Medium 3.5 per Mistral's own Legacy/Deprecated table (deprecated 2026-05-22, retirement 2026-07-31). Whole Magistral line retired; no successor reasoning model.
Deprecated — superseded by GPT-5.5. OpenAI's deprecations page now lists an API shutdown of 2026-12-11 for o3-2025-04-16 (announced 2026-06-11), reversing the prior 'no API change' position recorded below.
Deprecated — retired with no successor. OpenAI's deprecations page lists all sora-2 aliases and snapshots shutting down 2026-09-24 (announced 2026-03-24) with an empty replacement column. Recorded without superseded_by; the schema refinement was relaxed to permit successor-less retirements.
Updated OpenAI column from gpt-image-1 to gpt-image-2 (released 2026-04-21): category-leading text rendering, native 2K output, LLM-planned rendering, token-based pricing tiers.
Initial entry. Released 2026-06-09 as Anthropic's most capable widely released model — the first publicly available Mythos-class model, a new tier above Opus 4.8.
Initial entry. Released 2026-06-09 in limited availability via Project Glasswing, succeeding Claude Mythos Preview.
Repositioned as Anthropic's most capable Opus-tier model: Claude Fable 5 launched 2026-06-09 as a new tier above Opus. Opus 4.8 is not superseded and remains the standard-lineup flagship.
Verified current vs Anthropic docs. Corrected release date to 2026-02-17 (prior 2025-09-29 was the Sonnet 4.5 snapshot date).
Deprecated — superseded by Claude Opus 4.8; Opus 4.7 is now in Anthropic's legacy table.
Initial entry (2026-06-08 catalog refresh, Tier 3). Current flagship Opus, released 2026-05-28; supersedes Claude 4.7 Opus.
Updated from Embed v3 to embed-v4.0 (released 2025-04-15): multimodal, 128K context. The v3.0 models remain available but are no longer the flagship.
Updated from Rerank v3 to rerank-v4.0-pro (released 2025-12-11). (Interim: rerank-v3.5 2024-12-02.)
Updated to distil-large-v3.5 (trained on ~4× more data; faster than Whisper-large-v3-Turbo on long-form).
Updated from Multilingual v2 to Eleven v3 (GA 2026-02-02), 70+ languages. Multilingual v2 remains available but is no longer the flagship.
Updated from FLUX.1 [pro] to FLUX.2 [pro] (announced 2025-11-25), the current commercial flagship. FLUX.1 [pro] remains callable via the legacy /flux-pro endpoint.
Deprecated — Gemini 2.5 Deep Think is no longer a current API model; Deep Think now lives in the Gemini 3.x generation.
Initial entry (2026-06-08 catalog refresh, Tier 3). GA model released 2026-05-19 at Google I/O; supersedes the Gemini 3 Flash preview.
Verified current vs Google docs. Corrected release date to 2026-04-02 (Gemma 4 launch per Google blog; prior 2026-05-07 was an HF lastModified timestamp).
Deprecated — superseded by GPT-5.5 (OpenAI 'Deprecated' badge; retired from ChatGPT 2026-02-13; oldest snapshot API shutdown 2026-10-23).
Verified current. The base alias now points to the gpt-4o-mini-tts-2025-12-15 snapshot; the 2025-03-20 snapshot retires 2026-07-23.
Initial entry (2026-06-08 catalog refresh, Tier 3). OpenAI's hosted-API STT model; complements the open-weights Whisper entries.
Initial entry (2026-06-08 catalog refresh, Tier 3). Released 2026-02-05; OpenAI's flagship agentic coding model.
Initial entry (2026-06-08 catalog refresh, Tier 3). Released 2026-03-17; OpenAI's named replacement for o4-mini.
Initial entry (2026-06-08 catalog refresh, Tier 3). Flagship released 2026-04-23; the successor OpenAI positions above GPT-4o and the o-series.
Updated from Kling 2.0 to Kling 3.0 (released 2026-02-05): native audio, multimodal input, up to 15s clips. (Interim: 2.1, 2.5 Turbo, 2.6.)
Updated to Luma Ray3 (released 2025-09-18), the model behind the Dream Machine app. (Interim: Ray2 early 2025; later Ray3 Modify 2025-12-18, Ray3.14 2026-01-26.)
Updated from v6 to V8.1 (released 2026-04-30) per updates.midjourney.com. (Interim: v7 2025-04-03, v8.0 Alpha 2026-03-17.)
Verified current vs Mistral docs (mistral-large-2512). Corrected release date to 2025-12-02 (Mistral 3 family launch).
Repointed from parakeet-tdt-1.1b to current flagship parakeet-tdt-0.6b-v2 (English, 2025-05-01); added parakeet-tdt-0.6b-v3 (multilingual, 2025-08-14).
Verified current. o3 remains available in the API — OpenAI is retiring it from ChatGPT only (2026-08-26), with no API change. GPT-5.5 is the newer reasoning-capable flagship.
Verified current vs Mistral docs. Corrected release date to 2025-12-17 (announcement date).
Updated from Pika 2 to Pika 2.2 (released 2025-02-27): Pikaframes, 1080p, up to 10s. (A 'Pika 2.5' is referenced on the site but unconfirmed by a dated source.)
Repointed to active repo OHF-Voice/piper1-gpl (rhasspy/piper archived 2025-10-06; latest release v1.4.2, 2026-04-02).
Replaced fabricated 'Qwen 3.5 122B' entry (no such repo) with the real current open-weight flagship Qwen3.6-35B-A3B (released 2026-04-17). The real Qwen3.5 flagship was Qwen3.5-397B-A17B (2026-02-16).
Updated from Gen-3 to Gen-4.5 (released 2025-12-11) per Runway changelog. (Interim: Gen-4 / Gen-4 Turbo 2025-04.)
Updated original Sora to Sora 2 (released 2025-09-30). OpenAI is winding the product down — app closed (2026-04-26), API shutdown slated 2026-09-24; no current OpenAI video successor is generally available.
Updated from Stable Audio 2.5 to Stable Audio 3.0 (released 2026-05-20). Prior release date (2024-04-03) belonged to the 2.0 line.
Updated from v4 to v5.5 (released 2026-03-26) per suno.com/release-notes. (Interim: v4.5 2025-05, v5 2025-09-23.)
Updated to Allegro v1.5 (released 2025-03-18), the latest named model in Udio's official changelog. (Third-party 'v3.5/v4' claims are unverified.)
Updated from Voyage 3 to voyage-4-large (released 2026-01-15). The voyage-3 series is still accessible but Voyage marks it 'not recommended for new implementations'. (Interim: voyage-3.5 2025-05.)
Initial publication.
Initial publication.
Initial publication.
Initial publication. Snapshot reflects flagship model versions GPT-5.4, Claude Sonnet 4.6, and Gemini 3.1 Pro available in May 2026.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication. Pricing snapshot reflects May 2026 flagship and budget tiers.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Marked deprecated after verification that Codestral 25.08 is the current Codestral release.
Marked deprecated after verification that GPT Image 2 is now OpenAI's current image generation model.
Added after verification of the Gemini 3 family rollout in current Google model docs.
Added after verification of the Gemini 3 family rollout in current Google model docs.
Added after verification of the Gemini 3 family rollout in current Google model docs.
Added after verification of the Gemini 3 family rollout in current Google model docs.
Marked deprecated after verification that Gemini 3.1 Flash Live Preview is the current Live API successor.
Marked deprecated after OpenAI listed GPT-4.1 as the replacement in its deprecations guide.
Added after verification that GPT Image 2 is the current OpenAI image generation model.
Marked deprecated after verification that Mistral 7B no longer appears in Mistral's current model overview.
Initial entry. Sighted on Hugging Face under Mistral-Large-3-675B-Instruct-2512 family.
Marked deprecated after verification that OCR 3 is the current Mistral OCR offering.
Marked deprecated after verification that Mixtral 8x7B no longer appears in Mistral's current model overview.
Marked deprecated after verification that GPT-4o mini TTS is OpenAI's current speech generation model.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication. Reflects post-August-2025 vendor policy changes (Anthropic opt-in default, retention extensions) and pre-August-2026 EU AI Act enforcement timeline.
Initial publication.
Initial publication.
Initial publication. Snapshot reflects Cursor (Pro/Pro+/Ultra), GitHub Copilot (Pro/Pro+/Business/Enterprise), and Claude Code via Claude Pro/Max plans.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Initial publication.
Updated to recommend Whisper large-v3 over medium-en for English. Throughput numbers updated for new model.
Added Apple Silicon sizing notes; clarified break-even math.
Initial publication.