Cyberax AI Playbook
cyberax.com
Changelog · Catalog-wide

What changed, and when

Every update to every solution and model entry. Newest first.

Updates

2026-09-08
AI hallucinations explained

Model names refreshed to the September 2026 lineup. The benchmark figures deliberately keep the model they were measured on, because those tests have not been re-run on the newer models; they are now dated so the age is visible.

2026-09-08
AI translation services compared

Cost rows updated to Claude Sonnet 5 ($2/$10) and GPT-5.6 Terra ($2/$12), replacing Sonnet 4.6 and the deprecated GPT-4o. The Sonnet per-character estimate only fell to ~$3.50 rather than by the full price cut, because Claude 5's tokenizer produces roughly 30% more tokens for the same text.

2026-09-08
AI writing tools compared

Verified against claude.com/pricing. Corrected the Claude Team price — it is $20/seat annually, not $25 — and added the Team Premium tier at $100/seat. Updated the model-under-the-hood row from Sonnet 4.6 / Opus 4.7 to Sonnet 5 / Opus 5, noting that Opus requires a paid plan. Jasper, Copy.ai and Writer pricing was not re-verified this round.

2026-09-08
Auto-categorize support tickets by topic and urgency

Model names in the inference-cost row updated to Claude Sonnet 5 and GPT-5.6 Terra. GPT-4o was named here and is deprecated, superseded by GPT-5.5. The cost range is unchanged. Added referenced-model links for models this page already names: Claude Sonnet 5, GPT-5.6 Terra.

2026-09-08
Auto-tag and route inbound social DMs

Referenced-model links updated to the current lineup: Gemini 3.1 Flash-Lite to Gemini 3.5 Flash-Lite. Page content not otherwise re-verified.

2026-09-08
ChatGPT Team vs Claude Pro vs Microsoft Copilot for small business

Two corrections. Claude's context window is set by the model, not the plan: any paid plan gets 1M when chatting with Opus 5, Sonnet 5 or Fable 5.1, and older models give 500K or 200K. And Microsoft names no model at all, routing automatically per prompt. Flagship models updated to Claude Opus 5 and Sonnet 5.

2026-09-08
Contract review and clause extraction

Corrected the model prerequisite. It described Claude as '200k default, 1M on the higher tiers', which is backwards for the current generation: every Claude 5 model is 1M with no smaller variant to fall back to. The deprecated GPT-4o reference was replaced.

2026-09-08
CRM data hygiene at scale

Referenced-model links updated to the current lineup: Gemini 3.1 Flash-Lite to Gemini 3.5 Flash-Lite. Page content not otherwise re-verified.

2026-09-08
Cursor vs Copilot vs Claude Code for coding assistance

Model and plan rows re-verified against each vendor's own docs. Cursor does not offer GPT-6 Astra; GitHub Copilot has added a Max tier at $100/month; Claude Code defaults to Opus 5 on Max and above and Sonnet 5 on Pro. Claude Pro corrected to $17/month billed annually.

2026-09-08
Draft customer support replies that hold up to scrutiny

Recommended default models updated from Claude Sonnet 4.6 and GPT-4o to Claude Sonnet 5 and GPT-5.6 Terra, in both the prose and the cost row. GPT-4o is deprecated and two generations behind, so the old recommendation was steering readers at a model OpenAI has superseded. Added referenced-model links for models this page already names: Claude Sonnet 5, GPT-5.6 Terra.

2026-09-08
Document AI services compared

Updated the Claude Vision reference from Sonnet 4.6 / Opus 4.7 to Sonnet 5 / Opus 5. Repointed the referenced-models links to Mistral OCR 4.1 and the current Claude and GPT entries.

2026-09-08
Document classification at scale

Referenced-model links updated to the current lineup: Gemini 3.1 Flash-Lite to Gemini 3.5 Flash-Lite. Added referenced-model links for models this page already names: GPT-4.1, OpenAI text-embedding-3-large. Page content not otherwise re-verified.

2026-09-08
Email-to-task automation

Referenced-model links updated to the current lineup: Gemini 3.1 Flash-Lite to Gemini 3.5 Flash-Lite. Added referenced-model links for models this page already names: GPT-4.1. Page content not otherwise re-verified.

2026-09-08
Embeddings explained without math

Added referenced-model links for models this page already names: BGE-M3, Cohere Embed v4, OpenAI text-embedding-3-large, OpenAI text-embedding-3-small. Page content not otherwise re-verified.

2026-09-08
Extract structured data from PDFs at scale

Flagship-model cost row updated to Claude Sonnet 5 / GPT-5.6 Terra / Gemini 3.1, and the referenced models repointed to Mistral OCR 4.1 and the current Claude and GPT entries. The per-page cost range is unchanged.

2026-09-08
Federated search across your tools

Referenced-model links updated to the current lineup: Gemini 3.1 Flash-Lite to Gemini 3.5 Flash-Lite. Page content not otherwise re-verified.

2026-09-08
Find patterns in customer feedback

Referenced-model links updated to the current lineup: Claude 4.6 Sonnet to Claude Sonnet 5; GPT-4o to GPT-5.6 Terra. Added referenced-model links for models this page already names: OpenAI text-embedding-3-small. Page content not otherwise re-verified.

2026-09-08
Generate alt text and image descriptions at scale

Updated the per-image vision cost row to Claude Sonnet 5 / GPT-5.6 Terra / Gemini 3.1. The cost range is unchanged. Referenced-model links updated to the current lineup: GPT-4o to GPT-5.6 Terra; Claude 4.7 Opus to Claude Opus 5. Added referenced-model links for models this page already names: Claude Sonnet 5, GPT-5.5.

2026-09-08
GPT vs Claude vs Gemini for business writing

Pricing, context windows and model names refreshed to the September 2026 lineup. The quality rankings still come from the Q1 2026 blind evaluation of the versions current at the May 2026 snapshot; that has not been re-run, and the page now says so up front.

2026-09-08
Hook generation for short-form video

Referenced-model links updated to the current lineup: Gemini 3.1 Flash-Lite to Gemini 3.5 Flash-Lite. Page content not otherwise re-verified.

2026-09-08
Internal Q&A bot over company docs

Referenced-model links updated to the current lineup: Gemini 3.1 Flash-Lite to Gemini 3.5 Flash-Lite. Page content not otherwise re-verified.

2026-09-08
Lease and vendor renewal tracking

Referenced-model links updated to the current lineup: Gemini 3.1 Flash-Lite to Gemini 3.5 Flash-Lite. Page content not otherwise re-verified.

2026-09-08
When to run AI locally vs in the cloud

Frontier and open-weight model references refreshed to the September 2026 lineup. The open-weight row is now tiered by hardware, because naming DeepSeek V4 Pro flat in the 'Local' column implied a 1.6T model runs on the workstation budgeted two rows below.

2026-09-08
Open-source vs proprietary AI — practical tradeoffs

Updated the proprietary-flagship price stat to the September 2026 lineup (GPT-6 Astra, Claude Opus 5, Gemini 3.1 Pro). The range moved from $3–$15 to $2–$10 per million input tokens. Added referenced-model links for models this page already names: Claude Haiku 4.5, Claude Opus 5, Gemini 3.5 Flash-Lite, GPT-5.6 Luna, GPT-6 Astra.

2026-09-08
Build a private RAG with no third-party calls

The model-choice step now leads with Qwen3.8 27B, which is Apache 2.0 and fits the single-GPU pilot tier, instead of Llama 3.3 70B, which pushed readers toward far more expensive hardware than the job needs. The measured latency figures stay with the model they were measured on.

2026-09-08
Procurement RFP response comparison

Referenced-model links updated to the current lineup: Gemini 3.1 Flash-Lite to Gemini 3.5 Flash-Lite. Page content not otherwise re-verified.

2026-09-08
Prompt engineering patterns for content teams

Cost estimate moved from Claude Sonnet 4.6 to Sonnet 5 pricing. The range tightened only slightly (~$0.01-$0.035 from ~$0.01-$0.04) rather than falling by the full third the price cut implies, because Claude 5's tokenizer produces roughly 30% more tokens for the same text. Added referenced-model links for models this page already names: Claude Sonnet 5.

2026-09-08
Slack channel summaries that catch what matters

Corrected the model-API prerequisite from 'Claude (200k/1M), GPT (128k-272k)' to the current windows: Claude 5 is 1M on every model, GPT-5.6 and GPT-6 reach ~1.05M. last_verified deliberately not bumped: only this line was re-checked, not the whole page. Referenced-model links updated to the current lineup: Gemini 3.1 Flash-Lite to Gemini 3.5 Flash-Lite.

2026-09-08
Summarize long email threads

Per-summary cost recalculated at Claude Sonnet 5 pricing: about $0.03, down from $0.05. Claude's context figure corrected too, since any paid plan chatting with Opus 5, Sonnet 5 or Fable 5.1 gets the full 1M.

2026-09-08
Tokens, context windows, and what they cost

New section on what actually sets your context window: the model you pick, not the plan you pay for. Any paid Claude plan gets 1M on Opus 5, Sonnet 5 or Fable 5.1, while older models give 500K or 200K. Pricing refreshed to the September 2026 lineup and both worked cost examples recalculated.

2026-09-08
Triage inbound email at scale

Referenced-model links updated to the current lineup: Gemini 3.1 Flash-Lite to Gemini 3.5 Flash-Lite. Added referenced-model links for models this page already names: GPT-4.1. Page content not otherwise re-verified.

2026-09-08
Vector databases compared

Added referenced-model links for models this page already names: BGE-M3. Page content not otherwise re-verified.

2026-09-08
What "AI agents" actually are (and aren't)

Updated the agent-capability example from Claude Sonnet 4.5+/Opus 4.7 to Claude Sonnet 5 / Opus 5. Added referenced-model links for models this page already names: Claude Sonnet 5.

2026-09-08
What an LLM actually does for a business

Referenced-model links updated to the current lineup: GPT-4o to GPT-5.6 Terra; Claude 4.7 Opus to Claude Opus 5. Added referenced-model links for models this page already names: GPT-5.5. Page content not otherwise re-verified.

2026-09-08
BGE-M3 Model

Verified current. No bge-m4 exists; BGE-M3's successors are task-specific siblings rather than a new generation.

2026-09-08
BGE Reranker v2 Model

Verified current. No bge-reranker-v3 exists. The newer v2.5-gemma2-lightweight is a heavier variant under the restrictive Gemma licence, not a general replacement.

2026-09-08
Claude 4.6 Sonnet Model

Deprecated — superseded by Claude Sonnet 5 (released 2026-06-30), documented by Anthropic as a drop-in replacement. Sonnet 4.6 moves to the legacy table; retirement not sooner than 2027-02-17.

2026-09-08
Claude Fable 5 Model

Deprecated — superseded by Claude Fable 5.1 (released 2026-09-01). Fable 5 moves to Anthropic's legacy table; still available, retirement not sooner than 2027-06-09.

2026-09-08
Claude Fable 5.1 Model

Initial entry. Released 2026-09-01 as the successor to Claude Fable 5, at the same $10/$50 per MTok with cache reads cut to $0.25 per MTok. Adds adaptive thinking always on; drops support for forced tool use.

2026-09-08
Claude Haiku 4.5 Model

Verified current against Anthropic's models overview. Still a column in the current-lineup table; no Haiku 5 exists. Retirement not sooner than 2026-10-15.

2026-09-08
Claude Mythos 5 Model

Deprecated — superseded by Claude Mythos 5.1 (released 2026-09-01). Still in limited availability; not listed on Anthropic's deprecation table.

2026-09-08
Claude Mythos 5.1 Model

Initial entry. Released 2026-09-01 alongside Claude Fable 5.1, succeeding Claude Mythos 5. Remains invitation-only via Project Glasswing.

2026-09-08
Claude Opus 4.8 Model

Deprecated — superseded by Claude Opus 5 (released 2026-07-24). Opus 4.8 moves to Anthropic's legacy table; still available on all platforms, retirement not sooner than 2027-05-28.

2026-09-08
Claude Opus 5 Model

Initial entry. Released 2026-07-24 as the Opus-tier flagship, superseding Claude Opus 4.8. 1M-token context (default and maximum), $5/$25 per MTok, adaptive thinking on by default with a low/medium/high/xhigh/max effort ladder.

2026-09-08
Claude Sonnet 5 Model

Initial entry. Released 2026-06-30 as a drop-in replacement for Claude Sonnet 4.6. The $2/$10 per MTok launch pricing became the standard price — the increase to $3/$15 scheduled for 2026-09-01 was cancelled.

2026-09-08
Codestral 25.08 Model

Verified current against Mistral's models overview — the only model in the Code Models section, and not on the deprecation table.

2026-09-08
Cohere Embed v4 Model

Verified current against Cohere's models page. embed-v4.0 is still the newest Embed generation; everything below is v3.0/v2.0. No v5 exists.

2026-09-08
Cohere Rerank v4 Model

Verified current against Cohere's models page and the Rerank v4.0 changelog. rerank-v4.0-pro is top of the rerank list; rerank-v4.0-fast is a sibling tier, not a successor.

2026-09-08
Coqui XTTS v2 Model

Verified: the weights remain hosted, ungated and heavily used (7.2M downloads last month) despite Coqui the company being defunct — this entry's existing 'community maintenance only' weakness is accurate. Maintenance has moved to idiap/coqui-ai-TTS.

2026-09-08
DeepSeek V4 Flash Model

Corrected on three counts. Release date 2026-05-06 was a Hugging Face push timestamp, not a release — the official DeepSeek-V4-Flash-0731 shipped 2026-07-31 per DeepSeek's API changelog, superseding the preview this entry described. Licence corrected from 'DeepSeek License' to MIT, and the false 'restrictions on commercial use cases' weakness removed. Added parameter size and 1M context window.

2026-09-08
DeepSeek V4 Pro Model

Corrected on three counts. Release date 2026-05-06 was a Hugging Face push timestamp, not a release — the GA DeepSeek-V4-Pro-0813 shipped 2026-08-13 per DeepSeek's API changelog, superseding the preview this entry described. Licence corrected from 'DeepSeek License' to MIT, and the false 'restrictions on commercial use cases' weakness removed. Added parameter size and 1M context window.

2026-09-08
Devstral 2 Model

Deprecated — Mistral's deprecation table shows devstral-2512 deprecated 2026-05-22 and retired 2026-07-31, with Mistral Medium 3.5 as the stated alternative. The entry had been showing a retired model as current.

2026-09-08
Distil-Whisper Model

Verified current against the full distil-whisper org (12 models). The distil-large-v3.5 family is newest; no v4 exists. Worth knowing: distil-large-v3 still has roughly 90x the downloads of v3.5.

2026-09-08
ElevenLabs Eleven v3 Model

Dropped Turbo v2.5 from purpose and strengths — ElevenLabs' models page now marks Eleven Turbo v2.5 and Turbo v2 as deprecated, 'outclassed by Flash models', with eleven_flash_v2_5 as the replacement. Eleven v3 itself verified current; no v4 exists.

2026-09-08
FLUX.2 [pro] Model

Verified current against BFL's FLUX.2 launch post. No newer [pro] version exists; FLUX.2 [max] is a sibling tier, not a successor.

2026-09-08
FLUX.2 [dev] Model

Verified current against Black Forest Labs' docs, which state FLUX.2 remains fully supported for production. FLUX 3 exists but its Dev open-weight backbone has not shipped.

2026-09-08
Gemini 2.5 Flash Model

Verified current against Google's models and deprecations pages. Still GA and priced; no shutdown announced. Generations behind the Flash line but not superseded within its own 2.5 family.

2026-09-08
Gemini 2.5 Pro Model

Verified current against Google's models and deprecations pages. Still GA with no shutdown announced — and now the only GA (non-preview) Pro model, since Gemini 3.5 Pro was shelved.

2026-09-08
Gemini 3.1 Flash-Lite Model

Deprecated by Google — the deprecations page gives a 2027-05-07 shutdown and names Gemini 3.5 Flash-Lite (2026-07-21) as the replacement.

2026-09-08
Gemini 3.1 Flash Live Preview Model

Release date corrected to 2026-03-11 from Google's deprecations page. The previous 2026-05-07 was the models page's global 'last updated' stamp. Verified still current with no shutdown announced.

2026-09-08
Gemini 3.1 Pro Preview Model

Verified current. Still the top Pro model on Google's docs — Gemini 3.5 Pro was announced at I/O, missed three times and reportedly shelved to pretrain Gemini 4. No Pro model newer than this exists.

2026-09-08
Gemini 3.5 Flash Model

Deprecated — Google's models page now describes it as the legacy Flash model. Superseded by Gemini 3.8 Flash (2026-09-02), with 3.6 and 3.7 Flash released in between. No shutdown date announced.

2026-09-08
Gemini 3.5 Flash-Lite Model

Initial entry. Released 2026-07-21; named on Google's deprecations page as the replacement for Gemini 3.1 Flash-Lite (shutdown 2027-05-07). Context window left unset — Google publishes none.

2026-09-08
Gemini 3.8 Flash Model

Initial entry. Released 2026-09-02 per Google's Gemini API changelog as the head of the Flash line, superseding Gemini 3.7 Flash (2026-08-13) and 3.6 Flash (2026-07-21). Dates taken from the changelog and deprecations pages, not the models page (whose per-model stamps are a global page timestamp).

2026-09-08
Gemini 3 Flash Preview Model

Release date corrected to 2025-12-17 from Google's deprecations page; the previous 2026-05-07 was the models page's global 'last updated' stamp. Note Google names gemini-3.6-flash as the official replacement, though the catalog chains it through Gemini 3.5 Flash.

2026-09-08
Gemma 4 31B Model

Verified current against Google's Gemma pages. Gemma 4 (April 2026) is the newest generation and 31B the top core size. Gemma is absent from the Gemini API models page by design — that is a coverage gap, not a deprecation.

2026-09-08
GLM-5.3-Flash Model

Initial entry — first Z.ai model in the catalog. Released 2026-08-26, eight days after the flagship GLM-5.3 (2026-08-18). The Flash variant is catalogued rather than the flagship because it is plain MIT, while GLM-5.3 carries a custom licence requiring a Z.ai security review for MaaS operators above $10B revenue. Pricing recorded at list ($0.15/$0.50); a 50% promotional rate expires 2026-09-09. Release date taken from Z.ai's release notes, NOT from the arXiv paper linked on the model card — that paper (2026-02-17) is the shared GLM-5 base-model report, not this model's release.

2026-09-08
GPT-4o mini TTS Model

Verified current against OpenAI's models overview — still listed under speech.

2026-09-08
GPT-4o Transcribe Model

Deprecated by OpenAI — announced 2026-08-26, shutdown 2027-02-26, replaced by gpt-transcribe and gpt-live-transcribe.

2026-09-08
GPT-5.6 Luna Model

Pricing corrected to $0.20 input / $1.20 output per MTok — OpenAI cut Luna's rates about 80% on 2026-07-30; the entry carried the July launch prices.

2026-09-08
GPT-5.6 Sol Model

Pricing corrected to $4 input / $20 output per MTok — OpenAI cut Sol's rates on 2026-08-21 (promotional at least through 2026-11-21); the entry carried the July launch prices. Repositioned below GPT-6 Astra (released 2026-09-03).

2026-09-08
GPT-5.6 Terra Model

Pricing corrected to $2 input / $12 output per MTok — OpenAI cut Terra's rates roughly 20% on 2026-07-30; the entry carried the July launch prices.

2026-09-08
GPT-6 Astra Model

Initial entry. Released 2026-09-03 per OpenAI's API changelog as 'our most capable model, built for the hardest end-to-end work'. Specs and pricing verified against the gpt-6-astra model docs.

2026-09-08
GPT Image 2 Model

Verified current against OpenAI's models overview — still listed, snapshot gpt-image-2-2026-04-21.

2026-09-08
GPT Live Transcribe Model

Initial entry. Released 2026-07-28. Named alongside gpt-transcribe on OpenAI's deprecations page as a replacement for gpt-4o-transcribe (shutdown 2027-02-26). Context window and max output are not published by OpenAI.

2026-09-08
GPT Transcribe Model

Initial entry. Released 2026-07-28. Named on OpenAI's deprecations page as the replacement for gpt-4o-transcribe (shutdown 2027-02-26). Context window and max output are not published by OpenAI, so both are left unset.

2026-09-08
Grok 4.6 Model

Initial entry — first SpaceXAI model in the catalog. Released 2026-08-12. Maker recorded as SpaceXAI: that is the string on the copyright notice, the nav on every announcement page, the model docs' own prose, and the Arena leaderboard's provider column. The company was formerly xAI, and the API and SDK identifiers are still xai_sdk / XAI_API_KEY / api.x.ai. Reasoning controls run low / medium / high / xhigh with no way to disable.

2026-09-08
Grok Imagine Image 2.0 Model

Initial entry. Released 2026-08-07. SpaceXAI's announcement claims #2 worldwide on the Arena text-to-image and image-edit leaderboards as of the launch date; the live text-to-image leaderboard on 2026-09-04 shows it at #3, displaced by Microsoft's mai-image-2.6, with gpt-image-2 still #1. The entry records the current position rather than the launch claim. Current Image Edit standing was not re-checked.

2026-09-08
Kimi K3 Model

Initial entry — first Moonshot AI model in the catalog. Weights released 2026-07-27; the API launched 2026-07-17. The licence was read in full: contrary to widely repeated secondary reports, it contains NO revenue-sharing clause and no percentage. What it has is a negotiation gate — MaaS operators above $20M revenue over any rolling 12 months must agree separate terms first — plus an attribution requirement above 100M MAU or $20M monthly revenue. Internal use and access through Moonshot's own products are exempt.

2026-09-08
Kling 3.0 Model

Verified current against Kuaishou's own Kling 3.0 quickstart guide. Nothing newer than the 3.0 series is named; all 'Kling 4.0' material is forward-looking speculation.

2026-09-08
Llama 4 Maverick Model

Release date corrected to 2025-04-05, Meta's actual Llama 4 launch. The previous 2025-05-22 was the Hugging Face repo push timestamp. Still the newest Llama — there is no Llama 5; Meta's open-weight work moved to the Muse family (see Muse Glimmer 30B).

2026-09-08
Llama 4 Scout Model

Release date corrected to 2025-04-05, Meta's actual Llama 4 launch. The previous 2025-05-22 was the Hugging Face repo push timestamp. Still the newest Llama — there is no Llama 5; Meta's open-weight work moved to the Muse family (see Muse Glimmer 30B).

2026-09-08
Luma Ray3.2 Model

Updated from Luma Ray3 to Ray3.2 (released 2026-06-09), confirmed as the current model on lumalabs.ai/llm-info. Ray3.14 shipped 2026-01-26 in between. Note for future runs: docs.lumalabs.ai/docs/video-generation is a generation behind (it documents only the deprecated ray-2 line) — use lumalabs.ai/llm-info as the verification source instead. No Ray4 exists.

2026-09-08
Ministral 3 14B Model

Verified current against Mistral's models overview and model card; release date matches exactly.

2026-09-08
Mistral Large 3 Model

Verified current against Mistral's models overview and model card; release date matches exactly.

2026-09-08
Mistral Medium 3.5 Model

Release date corrected to 2026-04-28 from Mistral's own model card and changelog; the previous 2026-04-01 was inferred from the 26.04 version stamp. Verified current — Mistral now names it the alternative for both the retired Devstral 2 and the retired Magistral line, with reasoning exposed through the reasoning_effort parameter rather than a separate model.

2026-09-08
Mistral Small 4 Model

Verified current against Mistral's models overview and model card; release date matches exactly. Now described as a hybrid model unifying instruct, reasoning and coding.

2026-09-08
Muse Glimmer 30B Model

Initial entry. Released 2026-08-10 under the new meta-models Hugging Face org. Meta's open-weight work has moved off the Llama name — there is no Llama 5, and the closed-weight Muse Spark (2026-04-08) replaced Llama as Meta's flagship. Llama 3.1 8B is kept in the catalog because Llama 4's smallest model is 109B, leaving Muse Glimmer as the only current Meta option for a single consumer GPU.

2026-09-08
NVIDIA Parakeet Model

Verified current — no Parakeet v4 exists. Note nvidia/parakeet-unified-en-0.6b (2026) now beats this model both offline and dramatically in streaming, so the 'top-of-leaderboard English accuracy' claim needs a re-check against the current Open ASR leaderboard.

2026-09-08
OCR 3 Model

Superseded by OCR 4.1. Release date corrected to 2025-12-18 to match Mistral's model card. Still available and not on Mistral's deprecation table.

2026-09-08
OCR 4.1 Model

Initial entry. Released 2026-07-16 and generally available 2026-08-31 per Mistral's changelog; mistral-ocr-latest and mistral-ocr-4 now point at it. OCR 4.0 shipped nine weeks earlier and is skipped as an interim version.

2026-09-08
Phi-4 Model

Verified current against the model card — 14B, 16K context, MIT, no deprecation notice. Phi-5 does not exist.

2026-09-08
Pika 2.2 Model

Pika 2.5 confirmed as the current model — Pika's own developer docs list pika/pika-2.5/text-to-video and image-to-video with published pricing ($0.04/sec at 720p, $0.09/sec at 1080p, 5s text-to-video cap). Entry held at 2.2 because Pika publishes no release date for 2.5 anywhere, and this catalog does not guess release dates. Bump the title once a dated source appears.

2026-09-08
Piper Model

Verified actively maintained: OHF-Voice/piper1-gpl shipped v1.8.0 on 2026-09-04, with v1.5.0 through v1.7.0 across July and August 2026. The Open Home Foundation's README notes it is seeking maintainers.

2026-09-08
Qwen3.6 35B-A3B Model

Superseded by Qwen3.8 27B (August 2026). Note the filename is a legacy artifact of the corrected 2026-06-08 entry and no longer matches the title.

2026-09-08
Qwen3.8 27B Model

Initial entry. Qwen3.8 released August 2026, superseding Qwen3.6 (April 2026) with Qwen3.5 (February 2026) in between. Alibaba's model card gives only a month, so the release date is recorded as the first of the month rather than a guessed day. Benchmarks and licence taken from the model card, not from Hugging Face push timestamps.

2026-09-08
Real-ESRGAN Model

Verified current and not archived; last release v0.3.0 (2022-09-20). No successor exists — the README's GFPGAN/BasicSR links are complementary projects, not replacements. Canonical repo remains under the author's personal account; TencentARC/Real-ESRGAN returns 404.

2026-09-08
Runway Gen-4.5 Model

Verified current against Runway's model docs and changelog, which dates Gen-4.5 to 2025-12-11 — matching this entry exactly. No Gen-5 exists. Aggregator claims of a March 2026 launch contradict Runway's own changelog and were rejected.

2026-09-08
Stable Audio 3.0 Model

Verified current against Stability's news page, which dates Stable Audio 3.0 to 2026-05-20 — matching this entry exactly.

2026-09-08
Stable Diffusion 3.5 Model

Verified current. A full Hugging Face org enumeration by creation date shows no Stable Diffusion 4 repo, and Stability's own 2026 news page never mentions one — the widely circulated 'SD4, April 2026' claim is fabricated.

2026-09-08
Stable Diffusion x4 Upscaler Model

Verified current — no newer open-weights Stability upscaler exists in the full org listing.

2026-09-08
StarCoder2 15B Model

Verified current. No StarCoder3 exists. BigCode still ships evaluation datasets but has published no new model since 2024, so dormancy here is not staleness.

2026-09-08
Suno v5.5 Model

Verified current against suno.com and Suno's release notes, which confirm v5.5 shipped 2026-03-26 — matching this entry exactly. No v6 exists. Watch item: Suno has said current models will be deprecated when its Warner/BMG-licensed models ship later in 2026.

2026-09-08
SwinIR Model

Verified current as the canonical single-image restoration implementation. Repo is not archived; last real commit 2022-12-04. The authors' VRT/RVRT work is video restoration — adjacent, explicitly not a replacement.

2026-09-08
Topaz Gigapixel Model

Corrected on four counts. Renamed from 'Topaz Gigapixel AI' to 'Topaz Gigapixel' — Topaz dropped the 'AI' suffix. Release date moved from 2018-06-01 to 2025-09-16, the current product's date per the product page's schema markup; the 2018 date predated the current product entirely. Product URL updated from /gigapixel-ai to /gigapixel. Pricing corrected from a one-time licence to subscription-only. Version 1.3.3 shipped 2026-07-27.

2026-09-08
Udio Allegro v1.5 Model

Verified current against Udio's own help-centre changelog, which names Allegro v1.5 (2025-03-18) as the newest model — matching this entry exactly. Third-party claims of a 'Udio v4' or 'v3.5' have no official backing anywhere and were rejected again this run.

2026-09-08
Veo 3.1 Model

Release date corrected to 2025-10-15 from Google's deprecations page. The previous 2026-05-07 was the models page's global 'last updated' stamp, not a release date. Verified still current — no Veo 4 exists.

2026-09-08
Veo 3.1 Lite Model

Release date corrected to 2026-03-31 from Google's deprecations page. The previous 2026-05-07 was the models page's global 'last updated' stamp, not a release date. Verified still current.

2026-09-08
Vosk Model

Verified actively maintained — real feature commits through 2026-08-09, despite no tagged release since v0.3.50 (2024-04-22). The release gap is not staleness. No successor project exists.

2026-09-08
Voyage 4 Model

Verified current against Voyage's embeddings docs. voyage-4-large is still the best general-purpose retrieval model. Note voyage-context-4 is a different family (contextualised chunk embeddings), not a newer version of this one.

2026-07-10
DALL·E 3 Model

Link fix: OpenAI's image guide moved — /api/docs/guides/images now 404s. Repointed the provider and official_link to /api/docs/guides/image-generation.

2026-07-10
ElevenLabs Eleven v3 Model

Link fix: promoted the models overview (elevenlabs.io/docs/overview/models) to the primary docs link. The API reference was the only docs link, and it does not enumerate models — which made it a poor verification source. Kept as a secondary link.

2026-07-10
GPT-5.5 Model

Verified current after the GPT-5.6 launch — GPT-5.6 is an additional frontier tier, not a replacement. OpenAI's deprecations page does not list GPT-5.5, and still names it as the recommended successor for o3, GPT-5, and the chat-latest snapshots. Release date corrected 2026-04-23 → 2026-04-24 per the API changelog.

2026-07-10
GPT-5.6 Luna Model

Initial entry. GPT-5.6 family released 2026-07-09 per OpenAI's API changelog; specs and pricing verified against the gpt-5.6-luna model docs.

2026-07-10
GPT-5.6 Sol Model

Initial entry. GPT-5.6 family released 2026-07-09 per OpenAI's API changelog; specs and pricing verified against the gpt-5.6-sol model docs. API alias is `gpt-5.6`.

2026-07-10
GPT-5.6 Terra Model

Initial entry. GPT-5.6 family released 2026-07-09 per OpenAI's API changelog; specs and pricing verified against the gpt-5.6-terra model docs.

2026-07-10
Magistral Medium 1.2 Model

Deprecated — superseded by Mistral Medium 3.5 per Mistral's own Legacy/Deprecated table (deprecated 2026-05-22, retirement 2026-07-31). Whole Magistral line retired; no successor reasoning model.

2026-07-10
o3 Model

Deprecated — superseded by GPT-5.5. OpenAI's deprecations page now lists an API shutdown of 2026-12-11 for o3-2025-04-16 (announced 2026-06-11), reversing the prior 'no API change' position recorded below.

2026-07-10
Sora 2 Model

Deprecated — retired with no successor. OpenAI's deprecations page lists all sora-2 aliases and snapshots shutting down 2026-09-24 (announced 2026-03-24) with an empty replacement column. Recorded without superseded_by; the schema refinement was relaxed to permit successor-less retirements.

2026-06-10
Image generation models for business use

Updated OpenAI column from gpt-image-1 to gpt-image-2 (released 2026-04-21): category-leading text rendering, native 2K output, LLM-planned rendering, token-based pricing tiers.

2026-06-10
Claude Fable 5 Model

Initial entry. Released 2026-06-09 as Anthropic's most capable widely released model — the first publicly available Mythos-class model, a new tier above Opus 4.8.

2026-06-10
Claude Mythos 5 Model

Initial entry. Released 2026-06-09 in limited availability via Project Glasswing, succeeding Claude Mythos Preview.

2026-06-10
Claude Opus 4.8 Model

Repositioned as Anthropic's most capable Opus-tier model: Claude Fable 5 launched 2026-06-09 as a new tier above Opus. Opus 4.8 is not superseded and remains the standard-lineup flagship.

2026-06-08
Claude 4.6 Sonnet Model

Verified current vs Anthropic docs. Corrected release date to 2026-02-17 (prior 2025-09-29 was the Sonnet 4.5 snapshot date).

2026-06-08
Claude 4.7 Opus Model

Deprecated — superseded by Claude Opus 4.8; Opus 4.7 is now in Anthropic's legacy table.

2026-06-08
Claude Opus 4.8 Model

Initial entry (2026-06-08 catalog refresh, Tier 3). Current flagship Opus, released 2026-05-28; supersedes Claude 4.7 Opus.

2026-06-08
Cohere Embed v4 Model

Updated from Embed v3 to embed-v4.0 (released 2025-04-15): multimodal, 128K context. The v3.0 models remain available but are no longer the flagship.

2026-06-08
Cohere Rerank v4 Model

Updated from Rerank v3 to rerank-v4.0-pro (released 2025-12-11). (Interim: rerank-v3.5 2024-12-02.)

2026-06-08
Distil-Whisper Model

Updated to distil-large-v3.5 (trained on ~4× more data; faster than Whisper-large-v3-Turbo on long-form).

2026-06-08
ElevenLabs Eleven v3 Model

Updated from Multilingual v2 to Eleven v3 (GA 2026-02-02), 70+ languages. Multilingual v2 remains available but is no longer the flagship.

2026-06-08
FLUX.2 [pro] Model

Updated from FLUX.1 [pro] to FLUX.2 [pro] (announced 2025-11-25), the current commercial flagship. FLUX.1 [pro] remains callable via the legacy /flux-pro endpoint.

2026-06-08
Gemini 2.5 Deep Think Model

Deprecated — Gemini 2.5 Deep Think is no longer a current API model; Deep Think now lives in the Gemini 3.x generation.

2026-06-08
Gemini 3.5 Flash Model

Initial entry (2026-06-08 catalog refresh, Tier 3). GA model released 2026-05-19 at Google I/O; supersedes the Gemini 3 Flash preview.

2026-06-08
Gemma 4 31B Model

Verified current vs Google docs. Corrected release date to 2026-04-02 (Gemma 4 launch per Google blog; prior 2026-05-07 was an HF lastModified timestamp).

2026-06-08
GPT-4o Model

Deprecated — superseded by GPT-5.5 (OpenAI 'Deprecated' badge; retired from ChatGPT 2026-02-13; oldest snapshot API shutdown 2026-10-23).

2026-06-08
GPT-4o mini TTS Model

Verified current. The base alias now points to the gpt-4o-mini-tts-2025-12-15 snapshot; the 2025-03-20 snapshot retires 2026-07-23.

2026-06-08
GPT-4o Transcribe Model

Initial entry (2026-06-08 catalog refresh, Tier 3). OpenAI's hosted-API STT model; complements the open-weights Whisper entries.

2026-06-08
GPT-5.3-Codex Model

Initial entry (2026-06-08 catalog refresh, Tier 3). Released 2026-02-05; OpenAI's flagship agentic coding model.

2026-06-08
GPT-5.4 mini Model

Initial entry (2026-06-08 catalog refresh, Tier 3). Released 2026-03-17; OpenAI's named replacement for o4-mini.

2026-06-08
GPT-5.5 Model

Initial entry (2026-06-08 catalog refresh, Tier 3). Flagship released 2026-04-23; the successor OpenAI positions above GPT-4o and the o-series.

2026-06-08
Kling 3.0 Model

Updated from Kling 2.0 to Kling 3.0 (released 2026-02-05): native audio, multimodal input, up to 15s clips. (Interim: 2.1, 2.5 Turbo, 2.6.)

2026-06-08
Luma Ray3.2 Model

Updated to Luma Ray3 (released 2025-09-18), the model behind the Dream Machine app. (Interim: Ray2 early 2025; later Ray3 Modify 2025-12-18, Ray3.14 2026-01-26.)

2026-06-08
Midjourney V8.1 Model

Updated from v6 to V8.1 (released 2026-04-30) per updates.midjourney.com. (Interim: v7 2025-04-03, v8.0 Alpha 2026-03-17.)

2026-06-08
Mistral Large 3 Model

Verified current vs Mistral docs (mistral-large-2512). Corrected release date to 2025-12-02 (Mistral 3 family launch).

2026-06-08
NVIDIA Parakeet Model

Repointed from parakeet-tdt-1.1b to current flagship parakeet-tdt-0.6b-v2 (English, 2025-05-01); added parakeet-tdt-0.6b-v3 (multilingual, 2025-08-14).

2026-06-08
o3 Model

Verified current. o3 remains available in the API — OpenAI is retiring it from ChatGPT only (2026-08-26), with no API change. GPT-5.5 is the newer reasoning-capable flagship.

2026-06-08
o4-mini Model

Deprecated — superseded by GPT-5.4 mini; OpenAI API shutdown 2026-10-23.

2026-06-08
OCR 3 Model

Verified current vs Mistral docs. Corrected release date to 2025-12-17 (announcement date).

2026-06-08
Pika 2.2 Model

Updated from Pika 2 to Pika 2.2 (released 2025-02-27): Pikaframes, 1080p, up to 10s. (A 'Pika 2.5' is referenced on the site but unconfirmed by a dated source.)

2026-06-08
Piper Model

Repointed to active repo OHF-Voice/piper1-gpl (rhasspy/piper archived 2025-10-06; latest release v1.4.2, 2026-04-02).

2026-06-08
Qwen3.6 35B-A3B Model

Replaced fabricated 'Qwen 3.5 122B' entry (no such repo) with the real current open-weight flagship Qwen3.6-35B-A3B (released 2026-04-17). The real Qwen3.5 flagship was Qwen3.5-397B-A17B (2026-02-16).

2026-06-08
Qwen QwQ-32B Model

Deprecated — standalone QwQ reasoning superseded by Qwen3.x thinking mode.

2026-06-08
Runway Gen-4.5 Model

Updated from Gen-3 to Gen-4.5 (released 2025-12-11) per Runway changelog. (Interim: Gen-4 / Gen-4 Turbo 2025-04.)

2026-06-08
Sora 2 Model

Updated original Sora to Sora 2 (released 2025-09-30). OpenAI is winding the product down — app closed (2026-04-26), API shutdown slated 2026-09-24; no current OpenAI video successor is generally available.

2026-06-08
Stable Audio 3.0 Model

Updated from Stable Audio 2.5 to Stable Audio 3.0 (released 2026-05-20). Prior release date (2024-04-03) belonged to the 2.0 line.

2026-06-08
Suno v5.5 Model

Updated from v4 to v5.5 (released 2026-03-26) per suno.com/release-notes. (Interim: v4.5 2025-05, v5 2025-09-23.)

2026-06-08
Udio Allegro v1.5 Model

Updated to Allegro v1.5 (released 2025-03-18), the latest named model in Udio's official changelog. (Third-party 'v3.5/v4' claims are unverified.)

2026-06-08
Voyage 4 Model

Updated from Voyage 3 to voyage-4-large (released 2026-01-15). The voyage-3 series is still accessible but Voyage marks it 'not recommended for new implementations'. (Interim: voyage-3.5 2025-05.)

2026-05-10
BGE-M3 Model

Initial entry.

2026-05-10
Codestral Model

Marked deprecated after verification that Codestral 25.08 is the current Codestral release.

2026-05-10
Codestral 25.08 Model

Added after verification of Mistral's current coding lineup.

2026-05-10
DALL·E 3 Model

Marked deprecated after verification that GPT Image 2 is now OpenAI's current image generation model.

2026-05-10
DeepSeek V4 Flash Model

Initial entry. Sighted on Hugging Face (lastModified 2026-05-06).

2026-05-10
DeepSeek V4 Pro Model

Initial entry. Sighted on Hugging Face (lastModified 2026-05-06).

2026-05-10
Devstral 2 Model

Added after verification of Mistral's current frontier coding lineup.

2026-05-10
FLUX.2 [dev] Model

Initial entry. Sighted on Hugging Face (lastModified 2026-02-17).

2026-05-10
Gemini 3.1 Flash-Lite Model

Added after verification of the Gemini 3 family rollout in current Google model docs.

2026-05-10
Gemini 3.1 Pro Preview Model

Added after verification of the Gemini 3 family rollout in current Google model docs.

2026-05-10
Gemini 3 Flash Preview Model

Added after verification of the Gemini 3 family rollout in current Google model docs.

2026-05-10
Gemini Live Model

Marked deprecated after verification that Gemini 3.1 Flash Live Preview is the current Live API successor.

2026-05-10
Gemma 4 31B Model

Initial entry. Sighted on Hugging Face (lastModified 2026-05-07).

2026-05-10
GPT-4.1 Model

Added after verification of the GPT-4.1 family launch and current availability.

2026-05-10
GPT-4.1 mini Model

Added after verification of the GPT-4.1 family launch and current availability.

2026-05-10
GPT-4 Turbo Model

Marked deprecated after OpenAI listed GPT-4.1 as the replacement in its deprecations guide.

2026-05-10
GPT-4o Model

Initial entry.

2026-05-10
GPT-4o mini TTS Model

Added after verification of the current OpenAI speech generation lineup.

2026-05-10
GPT Image 2 Model

Added after verification that GPT Image 2 is the current OpenAI image generation model.

2026-05-10
Llama 4 Maverick Model

Initial entry. Sighted on Hugging Face (lastModified 2025-05-22).

2026-05-10
Llama 4 Scout Model

Initial entry. Sighted on Hugging Face (lastModified 2025-05-22).

2026-05-10
Magistral Medium 1.2 Model

Added after verification of Mistral's current reasoning lineup.

2026-05-10
Ministral 3 14B Model

Added after verification of Mistral's current local-deployment lineup.

2026-05-10
Mistral 7B Model

Marked deprecated after verification that Mistral 7B no longer appears in Mistral's current model overview.

2026-05-10
Mistral Large 3 Model

Initial entry. Sighted on Hugging Face under Mistral-Large-3-675B-Instruct-2512 family.

2026-05-10
Mistral Medium 3.5 Model

Added after verification of Mistral's current frontier multimodal lineup.

2026-05-10
Mistral OCR Model

Marked deprecated after verification that OCR 3 is the current Mistral OCR offering.

2026-05-10
Mistral Small 4 Model

Added after verification of Mistral's current frontier multimodal lineup.

2026-05-10
Mixtral 8x7B Model

Marked deprecated after verification that Mixtral 8x7B no longer appears in Mistral's current model overview.

2026-05-10
o3 Model

Initial entry.

2026-05-10
o4-mini Model

Initial entry.

2026-05-10
OCR 3 Model

Added after verification of Mistral's current OCR lineup.

2026-05-10
OpenAI TTS Model

Marked deprecated after verification that GPT-4o mini TTS is OpenAI's current speech generation model.

2026-05-10
Phi-4 Model

Initial entry.

2026-05-10
Pika 2.2 Model

Initial entry.

2026-05-10
Piper Model

Initial entry.

2026-05-10
Qwen3.6 35B-A3B Model

Initial entry. Sighted on Hugging Face (lastModified 2026-04-24).

2026-05-10
Sora 2 Model

Initial entry.

2026-05-10
SwinIR Model

Initial entry.

2026-05-10
Veo 3 Model

Initial entry.

2026-05-10
Veo 3.1 Model

Initial entry. Sighted in Gemini API docs on 2026-05-07.

2026-05-10
Veo 3.1 Lite Model

Initial entry. Sighted in Gemini API docs on 2026-05-07.

2026-05-10
Vosk Model

Initial entry.

2026-05-10
Voyage 4 Model

Initial entry.

2026-05-10
Whisper large-v2 Model

Backfilled lineage entry; v3 has been the recommended Whisper since 2023.

2026-05-07
AI privacy — what to watch for

Initial publication. Reflects post-August-2025 vendor policy changes (Anthropic opt-in default, retention extensions) and pre-August-2026 EU AI Act enforcement timeline.