Modelspublished

Microsoft Opens MAI-Image-2.6 in Foundry Preview and Adds a Lower-Cost Flash Model

The public preview gives Azure developers Microsoft’s own image-generation and editing models at two price and speed levels. The headline rankings and company speed figures are promising, but token billing and workload-specific reliability will determine the real production bargain.

By 3 min read
Microsoft Opens MAI-Image-2.6 in Foundry Preview and Adds a Lower-Cost Flash Model
Microsoft Opens MAI-Image-2.6 in Foundry Preview and Adds a Lower-Cost Flash Model

Listen to this story

The audio brief

About 1:26
0:001:26
Read transcript
Microsoft has opened public-preview access to its own image models in Microsoft Foundry, giving developers a direct choice between higher precision and lower-cost, faster image work. The flagship, MAI-Image-2.6, generates and edits images, accepts multiple reference images, and supports web grounding, aspect-ratio, format, and resolution controls. Microsoft says it ranked number two on Arena for both text-to-image generation and editing on September fourth. The new MAI-Image-2.6-Flash is aimed at latency-sensitive, high-volume applications instead. Microsoft describes Flash as production-quality, and says it generates images 2.8 times faster than GPT-Image-2-Medium, with 72 percent greater efficiency. Those are Microsoft’s comparisons, though, and the company cautions that real latency varies with prompts and system load. The pricing gap is substantial: image output starts at 38 dollars per million tokens for the flagship, versus 19 dollars for Flash. But Foundry does not charge a simple fixed price per picture. Text inputs, image inputs, and image outputs are metered separately, so resolution and reference images can change the bill. Flash also has lower input rates. The models first appeared in August, with broader access arriving September fourth. The key production test is now whether Flash preserves enough quality for a team’s own brand rules and workloads while delivering the promised speed and savings.

Story brief

3 key points

Microsoft is making its MAI-Image-2.6 available in public preview through Foundry, while adding MAI-Image-2.6-Flash for lower-cost, faster image generation and editing. The flagship costs $38 per million image-output tokens versus Flash’s $19; Flash also has lower input rates. The models accept multiple references and web grounding, but invoices vary with token usage and resolution. Public access on September 4 lets...

  1. 01

    Microsoft says MAI-Image-2.6 ranked No. 2 on Arena for text-to-image and editing on September 4.

  2. 02

    Flash claims 2.8× faster generation than GPT-Image-2-Medium and 72% greater efficiency; Microsoft cautions latency varies.

  3. 03

    Foundry meters text, image, and output tokens separately rather than charging a fixed price per generated image.

Microsoft has opened public-preview access to MAI-Image-2.6 through Microsoft Foundry and launched MAI-Image-2.6-Flash, a cheaper version aimed at applications that need fast, high-volume image work. The move gives developers a direct choice between the company’s flagship image quality and a model designed to reduce serving cost and response time.

Both models generate and edit images. They support multiple reference images, web grounding, and controls over aspect ratio, format, and resolution. Microsoft positions MAI-Image-2.6 as the precision option, while Flash is intended for latency-sensitive workloads, where a delay in returning an image or a high cost per request can shape the product experience.

One family, two operating targets

The flagship arrives with strong preference-benchmark positioning. Microsoft says MAI-Image-2.6 ranked No. 2 on Arena for both text-to-image generation and image editing on September 4. That is a quality signal, but it does not make Flash a like-for-like quality leader: Microsoft describes Flash as delivering comparable quality for production use rather than publishing the same ranking for the lower-cost model.

Speed claims use a different yardstick

Microsoft says Flash generates images 2.8 times faster than GPT-Image-2-Medium and delivers 72% greater efficiency. Those are company-supplied comparisons, not an independent measure of every production setup. Microsoft’s own launch material notes that latency can vary with the prompt and load, so the stated advantage is best read as a benchmarked performance claim rather than a guaranteed response time.

The price is metered by tokens, not pictures

The price gap is clearest on image output, but Foundry does not sell either model at a single fixed price per generated picture. It meters text inputs, image inputs, and image outputs separately. Image resolution, reference materials, and the output tokens a request consumes can all affect a developer’s invoice, making simple per-image estimates useful only under their stated assumptions.

  • MAI-Image-2.6 starts at $5 per million text-input tokens, $8 per million image-input tokens, and $38 per million image-output tokens.
  • Flash starts at $1.75 per million text-input tokens, $2.50 per million image-input tokens, and $19 per million image-output tokens.

Availability turns the score into a deployment decision

September 4 is the broader access point, not MAI-Image-2.6’s first appearance. Microsoft announced the model on August 10 and had placed it in its playground and a private Foundry preview by August 18. The public preview now lets developers test the more consequential questions on their own prompts: whether the flagship’s preference ranking holds for their brand rules and whether Flash keeps enough quality when speed and volume matter more.

Editorial analysis

Our Read

Microsoft is increasingly turning its MAI portfolio into a practical buying menu: one option for maximum output quality and another for speed and lower unit costs. This release makes that strategy visible in image generation, where MAI-Image-2.6 is now publicly testable beside Flash in Foundry. The next meaningful evidence will be whether customers can preserve brand consistency and editing quality while moving routine volume to Flash. Microsoft’s recent low-cost MAI-Transcribe-2 release points to a wider effort to compete on deployable economics, not only model scores.

Sources

  1. techcommunity.microsoft.comMAI-Image-2.6 and MAI-Image-2.6-Flash: Quality and speed at production scale | Microsoft Community Hub
  2. microsoft.aiPushing the quality-cost frontier with MAI-Image-2.6 | Microsoft AI