Release monitor / Verified records
AI Model Launch Tracker
Every consequential model release, structured.
A source-backed record of AI model launches, developers, availability, licenses, context windows, release types, and disclosed pricing.
Maintained dataset
Latest verified model launches signals.
Records update as sources arrive. Open any row to inspect captured facts, confidence, completeness, and supporting evidence.
Aug 17, 20263 sourcesGPT-5.6 SolOpenAIAPILicense not disclosedLimited release to a small group of customers; wider access planned later
Details
What we captured
OpenAI announced a limited API release of an Ultrafast mode for GPT-5.6 Sol, powered by Cerebras and capable of up to 750 output tokens per second.
- • OpenAI is testing an Ultrafast mode for GPT-5.6 Sol.
- • The mode is available to a small group of customers through the OpenAI API.
- • The mode is powered by Cerebras and can produce up to 750 output tokens per second.
Source evidence
Confidence 94% / Completeness 87%
Aug 11, 20261 sourceQwen3.8-27B-CRACK-GGUFJinho Jang (dealign.ai)262K contextLicense not disclosedHugging Face repository; downloadable GGUF files and a 0.9 GB vision projector for local use
Details
What we captured
Dealign.ai published Qwen3.8-27B-CRACK-GGUF, a 27B-parameter, safety-modified derivative of Qwen3.8-27B distributed as seven GGUF quantizations and a vision projector for local llama.cpp inference.
- • Repository named Qwen3.8-27B-CRACK-GGUF published and reproduced by dealign.ai
- • Described as a 27B-parameter derivative of Qwen3.8-27B with refusal behavior removed ("abliterated")
- • Repository metadata lists creation date as 2026-08-12
- • Distributed as seven GGUF quantizations (10.5 GB to 29.0 GB) and a 0.9 GB F16 vision projector
- • Recommended local build: 17.0 GB Q4_K_M quantization
- • Packaged for local multimodal inference with llama.cpp and provides OpenAI-compatible local server commands
Source evidence
Confidence 92% / Completeness 91%
Aug 10, 20261 sourceNemotron 3.5 LightningNvidia1M contextOpenMDW-1.1Available through Nvidia and model repositories including Hugging Face and ModelScope; NeMo Switchyard available on GitHub
Details
What we captured
Nvidia announced Nemotron 3.5 Lightning (30B MoE, 3B active per token), released Aug. 11, with open weights under OpenMDW-1.1 and availability via Nvidia, Hugging Face, and ModelScope; paired with open-source NeMo Switchyard routing library.
- • Model: Nemotron 3.5 Lightning — 30B parameters, mixture-of-experts
- • Active parameters per token: 3B
- • Released Aug. 11 (article text)
- • Context window: up to 1,000,000 tokens
- • License: OpenMDW-1.1, open weights/training data/recipes
- • Availability: Nvidia, Hugging Face, ModelScope; NeMo Switchyard on GitHub
Source evidence
Confidence 95% / Completeness 95%
Aug 9, 20263 sourcesMuse GlimmerMeta131.1K contextApache 2.0Consumer hardware with 24 GB or 32 GB of memory
Details
What we captured
Meta released the 30-billion-parameter multimodal Muse Glimmer model as open weights under Apache 2.0, with quantized configurations intended to run on 24 GB or 32 GB of memory and a 131,072-token context.
- • Meta announced the release on August 10, 2026.
- • The model has 30 billion parameters and is multimodal.
- • Meta says quantized configurations fit within 24 GB or 32 GB memory envelopes.
- • The release uses the Apache 2.0 license and supports a 131,072-token context.
Source evidence
Confidence 99% / Completeness 95%
Jul 15, 20261 sourceKimi K3Moonshot1M contextLicense requires companies operating a model-as-a-service business with > $20M total annual revenue (over 12 months) to enter a separate agreement with Moonshot before commercial use.Weights published for anyone to download; model can be run locally on user hardware
Details
What we captured
Moonshot introduced Kimi K3 on July 16, 2026 — an open-weights model described as 2.8 trillion parameters, multimodal, with up to a one-million-token context window; its license contains a commercial-use agreement requirement for high-revenue service providers.
- • Moonshot introduced Kimi K3 on July 16, 2026.
- • Kimi K3 described as 2.8 trillion parameters and native multimodal capabilities.
- • Kimi K3 supports a context window of up to 1,000,000 tokens.
- • Kimi K3 released with open weights published for public download.
- • Kimi K3 license clause: separate agreement required for model-as-a-service vendors with > $20M annual revenue.
Source evidence
Confidence 92% / Completeness 95%
Jul 7, 20261 sourceGPT-LiveOpenAIvoice model (full-duplex)License not disclosedRolling out to ChatGPT users globally (ChatGPT Voice); API availability planned/coming soon. (Update 2026-07-31: SynthID watermarking and API verification access added for supported audio.)
Details
What we captured
OpenAI announced GPT‑Live, a new generation of full‑duplex voice models (GPT‑Live‑1 and GPT‑Live‑1 mini) powering ChatGPT Voice, launched to ChatGPT users globally on July 8, 2026; OpenAI plans to bring the models to the API soon and at launch uses GPT‑5.5 in the background.
- • OpenAI announced GPT‑Live on July 8, 2026 as a new generation of voice models powering ChatGPT Voice.
- • GPT‑Live is built on a full‑duplex architecture and can listen and speak at the same time.
- • At launch GPT‑Live will use GPT‑5.5 in the background; OpenAI will update the background frontier model over time.
- • OpenAI is beginning to roll out GPT‑Live‑1 and GPT‑Live‑1 mini to ChatGPT users globally and plans to bring them to the API.
- • Update July 31, 2026: OpenAI added SynthID watermarking for supported audio, a public verification tool for provenance signals, and API access for verification.
Jun 8, 20262 sourcesClaude Fable 5AnthropicPublic release (Mythos-class); available via API and EnterpriseLicense not disclosedPublicly accessible via API and Enterprise; included in subscription plans through 2026-06-22 after which access requires purchase of usage credits; some sensitive-topic access restricted and Mythos 5 limited to vetted Project Glasswing participants
Details
What we captured
Anthropic publicly released Claude Fable 5 (a Mythos-class model) with topic-based safeguards; available via API and Enterprise with published per-million-token prices.
- • Article states Anthropic "publicly released Claude Fable 5, its first 'Mythos-class' model".
- • Fable 5 is publicly available while Mythos 5 remains restricted to a "small group of cyberdefenders" vetted through Project Glasswing.
- • Anthropic built topic-based safeguards in Fable 5 that funnel cybersecurity/biology/chemistry queries to Opus 4.8 and warn users; biology/chemistry classifier now applies to all such queries in Fable 5.
- • Anthropic says Fable 5 will be available to API and Enterprise users at $10 per million input tokens and $50 per million output tokens "starting today."
- • Subscription plans include access to Fable 5 through June 22, after which users must purchase "usage credits" to access the model.
Source evidence
Confidence 92% / Completeness 92%
Apr 6, 20262 sourcesClaude Mythos PreviewAnthropic1M contextLicense not disclosedrestricted access (restricted-access rollout)
Details
What we captured
Anthropic announced Claude Mythos Preview on April 7, 2026; the system card reports a 1,000,000-token context window, pricing listed as $25/$125 per million tokens, and a restricted-access rollout.
- • Article: "Anthropic announced Claude Mythos Preview on April 7, 2026."
- • Article: "The system card reports ... a 1M-token context window."
- • Article: "Mythos is priced at $25/$125 per million tokens — 5x above Opus ..."
- • Article: "The restricted-access rollout is consistent with a deeper pattern ..."
- • Article: "Mythos is Anthropic’s newest frontier model, internally codenamed 'Capybara.'"
Source evidence
Confidence 90% / Completeness 96%
Dec 15, 20251 sourceGPT Image 1.5OpenAIAPI and ChatGPT rolloutLicense not disclosedModel rolling out in ChatGPT for all users and available in the API as GPT Image 1.5; the new Images experience is rolling out to most ChatGPT users, with Business and Enterprise access coming later; global.
Details
What we captured
OpenAI announced a new ChatGPT Images experience and the GPT Image 1.5 image-generation model, rolling out December 16, 2025 in ChatGPT and available in the API.
- • Article published December 16, 2025 announcing the release.
- • "Today, we’re releasing a new version of ChatGPT Images... The new Images model is rolling out today in ChatGPT for all users, and is available in the API as GPT Image 1.5."
- • ChatGPT Images "makes precise edits while keeping details intact, and generates images up to 4x faster."
- • "Image inputs and outputs are now 20% cheaper in GPT Image 1.5 as compared to GPT Image 1."
- • The new Images experience in ChatGPT is "rolling out today for most users, with Business and Enterprise access coming later."
- • GPT Image 1.5 "delivers all the same improvements as ChatGPT Images: it’s stronger at image preservation and editing than GPT Image 1."
Dec 10, 20252 sourcesGPT-5.2OpenAI262.1K contextLicense not disclosedIn ChatGPT: GPT-5.2 Instant, Thinking, and Pro begin rolling out today to paid plans; In the API: available now to all developers.
Details
What we captured
OpenAI announced GPT-5.2, a new model series for professional knowledge work, with rollout in ChatGPT for paid plans and immediate availability in the API.
- • OpenAI published 'Introducing GPT‑5.2' on December 11, 2025.
- • Article states: 'In ChatGPT, GPT‑5.2 Instant, Thinking, and Pro will begin rolling out today, starting with paid plans.'
- • Article states: 'In the API, they are available now to all developers.'
- • Article notes GPT‑5.2 achieves near-100% accuracy on the 4-needle MRCR variant out to 256k tokens.
- • Article defines '256k represents 256 * 1,024 = 262,144 tokens.'
- • OpenAI describes GPT‑5.2 as 'the most capable model series yet for professional knowledge work' and highlights improvements in spreadsheets, presentations, code, vision, long-context understanding, and tool use.
Source evidence
Confidence 95% / Completeness 91%
