54 minutes agoModelsLaunchModelsAbliteration Removes Refusal Mechanisms for Offensive Cyber WorkThe derivative is built on Z.ai’s open-weight GLM-5.3, while Abliteration says its reasoning, coding and agent capabilities remain intact. The federal framework cited exempts open-source models, but this product is described as open-weight.3 min
7 hours agoModelsLaunchModelsMeta Rolls Out Muse Spark 1.3 for Longer Tasks, With Max Reasoning Still PendingThe update reaches Muse Code and Meta Model API with features meant to keep agent-style work on track. The highest reasoning setting remains unavailable until additional safety testing is complete.3 min
YesterdayModelsLaunchModelsGoogle May Ship Gemini 3.8 Flash 20 Days After 3.7, With Coding Edge UnprovenGoogle engineers reportedly preferred the internal model to an Anthropic Opus system in Jetski. Without the test design or public model materials, that result is a signal of internal usefulness rather than a broad coding verdict.3 min
YesterdayModelsLaunchModelsCursor Adds Claude Fable 5.1, Pairing a 73.4% Coding Score With Max-Effort CostsThe editor is betting that a model which keeps checking its work can carry harder coding jobs further. Its strongest published Cursor result, however, is at maximum effort—and an outside trial shows how sharply time and spend can rise at that setting.4 min
YesterdayModelsLaunchModelsWorld Labs Launches Atlas for 1440p Camera-Controlled Video, 3D Worlds and Robot ViewsThe new world model is designed to replace handoffs among video, reconstruction and simulation tools. Its performance evidence is company-run, and selected partners will get the first chance to test it on real work.3 min
YesterdayModelsLaunchModelsAnthropic Releases One Claude 5.1 Model in Two Access Tiers, Cuts Cache Reads 75%The product boundary is now access and safeguards: the general model is paired with a restricted version for cyber and life-sciences work. Anthropic estimates lower cache-read prices will cut typical token-billed costs by about 25%.2 min
YesterdayModelsLaunchModelsMeta Introduces Muse Voice Transcribe for Live Speaker Labels in 25 Validated LanguagesThe model puts transcription, speaker identification and speech-end detection into one streaming process, but Meta has not described how people or developers will access it.3 min
YesterdayModelsLaunchModelsGoogle Adds Agentic Video Understanding to Gemini, Claiming 88% Lower Token UseThe API feature shifts Gemini from fixed-rate video sampling to targeted inspection of frames, audio and transcripts. Its value for long recordings will depend on whether Google’s reported efficiency and accuracy gains carry into production workloads.3 min
YesterdayModelsLaunchModelsFlower Labs Starts Endeavor 1.0 Preview With a Private Deployment OptionThe model’s differentiator is not just Flower’s frontier-performance claim. Customers can run it as a managed service or place selected workloads inside their own infrastructure, though the launch remains limited to early users.2 min
Aug 31, 2026ModelsLaunchModelsSkild AI’s S1 Uses One Human Video to Prompt Robots Through 10-Minute TasksThe model’s central bet is that a demonstration can become immediate instruction rather than another task-specific training run. The harder question is whether that capability can turn into dependable deployments.3 min
Aug 31, 2026ModelsLaunchModelsGoogle’s TimesFM-3 Adds Known Future Events to Zero-Shot ForecastingThe 330-million-parameter model can combine related data streams with inputs such as promotion calendars and weather forecasts. Google reports leading results on three public benchmarks; performance on an organization’s own data remains the practical open question.3 min
Aug 27, 2026ModelsLaunchModelsAmap Releases ABot-Recon, a 12-Frame 3D Mapper for 10,000-Frame Video RunsThe model’s fixed local context could reduce the memory burden of long reconstruction jobs, but its headline speed is a data-center benchmark and its weights are not cleared for commercial deployment.3 min
Aug 27, 2026ModelsLaunchModelsMidjourney Opens V8.2 Image-Editing Tests With Up to Four ReferencesThe test brings instruction-led edits, canvas expansion, and multi-image generation into one workflow, while Midjourney asks users to identify failures and help reshape the interface.2 min
Aug 27, 2026ModelsLaunchModelsCohere’s Parse 5 Turns Enterprise Documents Into Markdown for $1.50 per 1,000 PagesThe compact vision model is built for high-volume retrieval and document-processing pipelines, but its headline benchmark result leaves out chart and visual-grounding tests.3 min
Aug 27, 2026ModelsLaunchModelsGemini Omni 1.1 Flash Adds 40-Second Scene Extensions and 4K Finishing ControlsGoogle is combining continuity controls, lower-cost previews and high-resolution finishing in one video workflow, though its extension and reference windows remain short.3 min
Aug 26, 2026ModelsLaunchModelsZ.ai Says 100,000 Chinese Chips Serve GLM-5.3-Flash; Shares Rise More Than 8%The low-cost model is a test of whether China-made hardware can support a public AI service at scale, but Z.ai has not named the chipmakers behind the system.3 min
Aug 26, 2026ModelsLaunchModelsxAI Puts Grok 4.6 on Microsoft Foundry, Extending a Two-Week Cloud RolloutAzure customers can now evaluate xAI’s model through managed endpoints and governance controls, but production terms for the public preview remain undefined.2 min
Aug 26, 2026ModelsLaunchModelsGoogle Puts Gemini 3.5 Transcribe in Preview With Live and Recorded-Audio APIsThe release turns speech recognition into a Gemini product layer for developers and Google surfaces, but a product-specific price remains unavailable for teams weighing production use.3 min