Loading page…
Loading page…
Release monitor / Verified records
Every consequential model release, structured.
A source-backed record of AI model launches, developers, availability, licenses, context windows, release types, and disclosed pricing.
Maintained dataset
Records update as sources arrive. Open any row to inspect captured facts, confidence, completeness, and supporting evidence.
Officially released; deployable on domestically produced AI chips
DetailsWhat we captured
Zhipu AI released GLM-5.3-Flash, an open-source MIT-licensed native multimodal model with a 1-million-token context window and 300 billion parameters. The company says it is served by more than 100,000 domestically produced chips.
Source evidence
Confidence 99% / Completeness 100%
Z.ai API; weights available on Hugging Face
DetailsWhat we captured
Z.ai released GLM-5.3-Flash with open weights, MIT licensing, a one-million-token context window, and API availability. The model is positioned as a lower-cost alternative to GLM-5.3 and was reportedly served on Chinese AI chips.
Consumer app and developer API
DetailsWhat we captured
Leiolai launched leiolai-1 alongside a consumer app and developer API. The model runs inference across users' devices and offers an 11-million-token context window.
Source evidence
Confidence 99% / Completeness 96%
API access
DetailsWhat we captured
Alibaba’s Qwen released Qwen3.8-Flash with a 262,144-token default context window, expandable to 1 million tokens, lower stated training costs, and API pricing of 1 yuan per million input tokens and 3 yuan per million output tokens.
Microsoft Foundry for Azure customers
DetailsWhat we captured
xAI made Grok 4.6 available to Azure customers through Microsoft Foundry in public preview, with a 500,000-token context window and published pricing of $2 per million input tokens and $6 per million output tokens.
QwenCloud managed API
DetailsWhat we captured
Alibaba made Qwen3.8-Flash available through the managed QwenCloud API at published rates, with a default 1-million-token context and OpenAI- and Anthropic-compatible interfaces.
Source evidence
Confidence 99% / Completeness 96%
Weights on Hugging Face and ModelScope; production Qwen3.8-Flash through QwenCloud; API expected shortly
DetailsWhat we captured
Alibaba’s Qwen team released Qwen3.8-Flash-Next, an open-weight multimodal mixture-of-experts architecture preview for Qwen4, with production access offered as Qwen3.8-Flash through QwenCloud.
Hugging Face, Ollama, GitHub, and other platforms
DetailsWhat we captured
IBM released Granite 4.2 in 3B, 8B, and 30B sizes under Apache 2.0, with up to 512,000-token context and agentic capabilities in the 8B and 30B variants.
Initially available in CoCounsel Legal’s Tabular Analysis feature; smaller version planned for Hugging Face
DetailsWhat we captured
Thomson Reuters launched Thomson, a Qwen-based legal language model trained with the company’s proprietary content and domain expertise. It is initially being deployed for document-review tasks in CoCounsel Legal, with a smaller open-weight, non-commercial version planned for Hugging Face.
Availability not disclosed
DetailsWhat we captured
Thomson Reuters launched Thomson, including Thomson-1.0-Large, a proprietary legal model specialized from open-weight foundations with Thomson Reuters' legal and news content, expert data, and evaluations.
Source evidence
Confidence 99% / Completeness 79%
Confidence 98% / Completeness 100%
Confidence 99% / Completeness 96%
Confidence 91% / Completeness 96%
Confidence 99% / Completeness 96%
Confidence 99% / Completeness 95%
Confidence 99% / Completeness 91%