Copilot's MAI-Code-1.1-Flash goes GA: 73% cheaper, adds vision

Microsoft's small coder hits every Copilot editor at a 0.25x multiplier; Ollama local models and cross-session memory also land in Copilot for JetBrains.

Nowline AUG 12 6:00 AM banner

Top AI stories from the last hour

Top AI stories from the last hour

Copy markdown

  • 73% cheaper - and it can finally see

    MAI-Code-1.1-Flash is generally available across Copilot CLI, VS Code, Visual Studio, JetBrains, Xcode, Eclipse, GitHub Mobile and github.com. Microsoft's small coder costs 73% less than the model it replaces, adds native image understanding, and bills at just a 0.25x premium multiplier for usage-based subscribers.

  • Mark the calendar: 1-Flash retires Sept 10

    GitHub will deprecate the older MAI-Code-1-Flash across all Copilot experiences on September 10, 2026. Repoint any pinned model configs to 1.1-Flash now; Enterprise and Business admins may first need to enable it via model policies.

  • Run local models inside Copilot for JetBrains

    The latest JetBrains plugin adds Ollama as a bring-your-own-key provider, so Copilot Chat can point at models running on your own machine. That is a real unlock for offline, air-gapped, or cost-sensitive work you could not route through the cloud before.

  • Copilot Memory lands in JetBrains too

    A new toggle in the Copilot settings portal lets Copilot retain project details and preferences across chat sessions, so you stop re-explaining your stack every conversation.

  • See exactly what is burning your AI credits

    The downloadable usage report now itemizes each model by input, output, cache-read and cache-write tokens next to the credits they consumed - on Business, Enterprise and individual plans. A surprise bill is finally explainable, and shrinkable.