Microsoft Corporation

08/11/2026 | Press release | Distributed by Public on 08/11/2026 15:22

MAI-Code-1.1 Flash: Better, faster, at a quarter of the cost

MAI-Code-1.1-Flash produces higher quality code, at 25% greater token efficiency, and at a quarter of the cost compared to the model we launched in June at Microsoft Build. This small, efficient, coding workhorse is now in production in GitHub Copilot.

We learned from developer feedback that CLI tasks and .NET performance mattered, so that's where we focused. The result: a 22% improvement on Terminal-Bench 2.1 in GitHub Copilot CLI and a 15% improvement on .NET tasks .

Benchmarks are useful guides but production is where the rubber meets the road. Most importantly, code survival rose 4% and return visits increased 9% .

1.1 is also dramatically more efficient. In GitHub Copilot tokens stream 25% faster and the model uses 25% fewer tokens to complete a task. That means faster answers, less waiting, and more useful work from every token-not simply a bigger model with a bigger bill.

Better training and serving efficiency let us offer a stronger, faster model at one quarter of the price of 1.0-and pass those savings reliably to customers. We achieved this by optimizing for real-world use across more than hundreds of thousands of reinforcement-learning environments in GitHub Copilot.

The loop is simple: ship, learn, improve, repeat. That's the MAI hill climbing machine.

Try MAI-Code-1.1-Flash today in GitHub Copilot, then tell us what needs improving by opening an issue here.

Microsoft Corporation published this content on August 11, 2026, and is solely responsible for the information contained herein. Distributed via Public Technologies (PUBT), unedited and unaltered, on August 11, 2026 at 21:23 UTC. If you believe the information included in the content is inaccurate or outdated and requires editing or removal, please contact us at [email protected]