Release monitor / Verified records

AI Model Launch Tracker

Every consequential model release, structured.

A source-backed record of AI model launches, developers, availability, licenses, context windows, release types, and disclosed pricing.

Release trajectoryModels entering the market
11 verified
Each launch rises from the release timeline. Taller signals indicate larger disclosed context windows.Hover or focus to inspect
Verified records11
Source evidence18
Named participants17
Named participants17

Maintained dataset

Latest verified model launches signals.

Records update as sources arrive. Open any row to inspect captured facts, confidence, completeness, and supporting evidence.

Showing 1-10 of 11 verified records
Aug 17, 20263 sources
GPT-5.6 SolOpenAI
APILicense not disclosed

Limited release to a small group of customers; wider access planned later

Details

What we captured

OpenAI announced a limited API release of an Ultrafast mode for GPT-5.6 Sol, powered by Cerebras and capable of up to 750 output tokens per second.

  • OpenAI is testing an Ultrafast mode for GPT-5.6 Sol.
  • The mode is available to a small group of customers through the OpenAI API.
  • The mode is powered by Cerebras and can produce up to 750 output tokens per second.
Aug 11, 20261 source
Qwen3.8-27B-CRACK-GGUFJinho Jang (dealign.ai)
262K contextLicense not disclosed

Hugging Face repository; downloadable GGUF files and a 0.9 GB vision projector for local use

Details

What we captured

Dealign.ai published Qwen3.8-27B-CRACK-GGUF, a 27B-parameter, safety-modified derivative of Qwen3.8-27B distributed as seven GGUF quantizations and a vision projector for local llama.cpp inference.

  • Repository named Qwen3.8-27B-CRACK-GGUF published and reproduced by dealign.ai
  • Described as a 27B-parameter derivative of Qwen3.8-27B with refusal behavior removed ("abliterated")
  • Repository metadata lists creation date as 2026-08-12
  • Distributed as seven GGUF quantizations (10.5 GB to 29.0 GB) and a 0.9 GB F16 vision projector
  • Recommended local build: 17.0 GB Q4_K_M quantization
  • Packaged for local multimodal inference with llama.cpp and provides OpenAI-compatible local server commands
Aug 10, 20261 source
Nemotron 3.5 LightningNvidia
1M contextOpenMDW-1.1

Available through Nvidia and model repositories including Hugging Face and ModelScope; NeMo Switchyard available on GitHub

Details

What we captured

Nvidia announced Nemotron 3.5 Lightning (30B MoE, 3B active per token), released Aug. 11, with open weights under OpenMDW-1.1 and availability via Nvidia, Hugging Face, and ModelScope; paired with open-source NeMo Switchyard routing library.

  • Model: Nemotron 3.5 Lightning — 30B parameters, mixture-of-experts
  • Active parameters per token: 3B
  • Released Aug. 11 (article text)
  • Context window: up to 1,000,000 tokens
  • License: OpenMDW-1.1, open weights/training data/recipes
  • Availability: Nvidia, Hugging Face, ModelScope; NeMo Switchyard on GitHub
Aug 9, 20263 sources
Muse GlimmerMeta
131.1K contextApache 2.0

Consumer hardware with 24 GB or 32 GB of memory

Details

What we captured

Meta released the 30-billion-parameter multimodal Muse Glimmer model as open weights under Apache 2.0, with quantized configurations intended to run on 24 GB or 32 GB of memory and a 131,072-token context.

  • Meta announced the release on August 10, 2026.
  • The model has 30 billion parameters and is multimodal.
  • Meta says quantized configurations fit within 24 GB or 32 GB memory envelopes.
  • The release uses the Apache 2.0 license and supports a 131,072-token context.
Jul 15, 20261 source
Kimi K3Moonshot
1M contextLicense requires companies operating a model-as-a-service business with > $20M total annual revenue (over 12 months) to enter a separate agreement with Moonshot before commercial use.

Weights published for anyone to download; model can be run locally on user hardware

Details

What we captured

Moonshot introduced Kimi K3 on July 16, 2026 — an open-weights model described as 2.8 trillion parameters, multimodal, with up to a one-million-token context window; its license contains a commercial-use agreement requirement for high-revenue service providers.

  • Moonshot introduced Kimi K3 on July 16, 2026.
  • Kimi K3 described as 2.8 trillion parameters and native multimodal capabilities.
  • Kimi K3 supports a context window of up to 1,000,000 tokens.
  • Kimi K3 released with open weights published for public download.
  • Kimi K3 license clause: separate agreement required for model-as-a-service vendors with > $20M annual revenue.
Jul 7, 20261 source
GPT-LiveOpenAI
voice model (full-duplex)License not disclosed

Rolling out to ChatGPT users globally (ChatGPT Voice); API availability planned/coming soon. (Update 2026-07-31: SynthID watermarking and API verification access added for supported audio.)

Details

What we captured

OpenAI announced GPT‑Live, a new generation of full‑duplex voice models (GPT‑Live‑1 and GPT‑Live‑1 mini) powering ChatGPT Voice, launched to ChatGPT users globally on July 8, 2026; OpenAI plans to bring the models to the API soon and at launch uses GPT‑5.5 in the background.

  • OpenAI announced GPT‑Live on July 8, 2026 as a new generation of voice models powering ChatGPT Voice.
  • GPT‑Live is built on a full‑duplex architecture and can listen and speak at the same time.
  • At launch GPT‑Live will use GPT‑5.5 in the background; OpenAI will update the background frontier model over time.
  • OpenAI is beginning to roll out GPT‑Live‑1 and GPT‑Live‑1 mini to ChatGPT users globally and plans to bring them to the API.
  • Update July 31, 2026: OpenAI added SynthID watermarking for supported audio, a public verification tool for provenance signals, and API access for verification.

Source evidence

Confidence 95% / Completeness 87%

Jun 8, 20262 sources
Claude Fable 5Anthropic
Public release (Mythos-class); available via API and EnterpriseLicense not disclosed

Publicly accessible via API and Enterprise; included in subscription plans through 2026-06-22 after which access requires purchase of usage credits; some sensitive-topic access restricted and Mythos 5 limited to vetted Project Glasswing participants

Details

What we captured

Anthropic publicly released Claude Fable 5 (a Mythos-class model) with topic-based safeguards; available via API and Enterprise with published per-million-token prices.

  • Article states Anthropic "publicly released Claude Fable 5, its first 'Mythos-class' model".
  • Fable 5 is publicly available while Mythos 5 remains restricted to a "small group of cyberdefenders" vetted through Project Glasswing.
  • Anthropic built topic-based safeguards in Fable 5 that funnel cybersecurity/biology/chemistry queries to Opus 4.8 and warn users; biology/chemistry classifier now applies to all such queries in Fable 5.
  • Anthropic says Fable 5 will be available to API and Enterprise users at $10 per million input tokens and $50 per million output tokens "starting today."
  • Subscription plans include access to Fable 5 through June 22, after which users must purchase "usage credits" to access the model.
Apr 6, 20262 sources
Claude Mythos PreviewAnthropic
1M contextLicense not disclosed

restricted access (restricted-access rollout)

Details

What we captured

Anthropic announced Claude Mythos Preview on April 7, 2026; the system card reports a 1,000,000-token context window, pricing listed as $25/$125 per million tokens, and a restricted-access rollout.

  • Article: "Anthropic announced Claude Mythos Preview on April 7, 2026."
  • Article: "The system card reports ... a 1M-token context window."
  • Article: "Mythos is priced at $25/$125 per million tokens — 5x above Opus ..."
  • Article: "The restricted-access rollout is consistent with a deeper pattern ..."
  • Article: "Mythos is Anthropic’s newest frontier model, internally codenamed 'Capybara.'"
Dec 15, 20251 source
GPT Image 1.5OpenAI
API and ChatGPT rolloutLicense not disclosed

Model rolling out in ChatGPT for all users and available in the API as GPT Image 1.5; the new Images experience is rolling out to most ChatGPT users, with Business and Enterprise access coming later; global.

Details

What we captured

OpenAI announced a new ChatGPT Images experience and the GPT Image 1.5 image-generation model, rolling out December 16, 2025 in ChatGPT and available in the API.

  • Article published December 16, 2025 announcing the release.
  • "Today, we’re releasing a new version of ChatGPT Images... The new Images model is rolling out today in ChatGPT for all users, and is available in the API as GPT Image 1.5."
  • ChatGPT Images "makes precise edits while keeping details intact, and generates images up to 4x faster."
  • "Image inputs and outputs are now 20% cheaper in GPT Image 1.5 as compared to GPT Image 1."
  • The new Images experience in ChatGPT is "rolling out today for most users, with Business and Enterprise access coming later."
  • GPT Image 1.5 "delivers all the same improvements as ChatGPT Images: it’s stronger at image preservation and editing than GPT Image 1."

Source evidence

Confidence 95% / Completeness 87%

Dec 10, 20252 sources
GPT-5.2OpenAI
262.1K contextLicense not disclosed

In ChatGPT: GPT-5.2 Instant, Thinking, and Pro begin rolling out today to paid plans; In the API: available now to all developers.

Details

What we captured

OpenAI announced GPT-5.2, a new model series for professional knowledge work, with rollout in ChatGPT for paid plans and immediate availability in the API.

  • OpenAI published 'Introducing GPT‑5.2' on December 11, 2025.
  • Article states: 'In ChatGPT, GPT‑5.2 Instant, Thinking, and Pro will begin rolling out today, starting with paid plans.'
  • Article states: 'In the API, they are available now to all developers.'
  • Article notes GPT‑5.2 achieves near-100% accuracy on the 4-needle MRCR variant out to 256k tokens.
  • Article defines '256k represents 256 * 1,024 = 262,144 tokens.'
  • OpenAI describes GPT‑5.2 as 'the most capable model series yet for professional knowledge work' and highlights improvements in spreadsheets, presentations, code, vision, long-context understanding, and tool use.