Generative AI & Microsoft.
As we see it.

Over Time

Microsoft is being pushed to own the stack while trust and cost tighten the terms of AI use

This week suggests Microsoft is trying to become less dependent on any one model supplier while becoming more useful as the place where enterprise AI is governed, deployed and paid for. Microsoft launched MAI image, voice and cyber models in Foundry and began moving workloads from OpenAI models to its own MAI models in products including Bing Image Creator, PowerPoint image features, OneDrive image editing, Dynamics 365 Contact Center and Azure Voice Live. It also added Anthropic’s Claude Opus 5 to both Microsoft 365 Copilot and Microsoft Foundry, while Manulife signed an expanded five-year partnership covering Microsoft 365 E7 Frontier Suite, Agent 365 governance and Copilot for more than 30,000 employees. This appears to extend last week’s line that competition is shifting from a single chatbot or base model toward the full stack of model choice, governance, memory, security and workflow deployment. It also suggests that for Microsoft, control of the customer environment may matter more than exclusive control of the underlying model.

The same week also showed why that stack is being built around constraints rather than around novelty alone. SAP decided to restrict AI token spending after internal use drove costs higher, and Alibaba Cloud described an agent optimisation aimed at cutting context-related token use. OpenAI launched Presence for governed enterprise agents, while Microsoft released a Cosmos DB-based memory provider for its Agent Framework and AWS updated AgentCore Gateway to the latest MCP specification with enterprise authorization and lifecycle controls. This likely means the next phase of enterprise AI buying will turn less on whether a company wants agents at all and more on whether those agents can be run within budgets, permissions and audit rules.

Trust pressure is rising at the same time. uniVersa reported that an OpenAI-linked crawler reached personal customer data during an IT migration, 404 Media reported that public Claude share links were being indexed by Google, Codeberg changed its terms to block AI training uses and heavily autogenerated projects, and the EU AI Omnibus entered into force. NIST also launched a sequestered evaluation programme, and the Reserve Bank of India issued draft model-risk guidance for regulated financial firms. We read this as evidence that AI is moving deeper into normal institutional control systems: privacy incidents, provenance signals, blind testing and formal risk management are no longer side issues, but part of the conditions under which deployment is allowed.

Underneath all of this, the infrastructure build-out still points in the same direction. Alphabet reported faster Google Cloud growth tied to AI infrastructure and enterprise AI demand, AMD launched its Helios rack-scale system and announced an Anthropic partnership for future GPU deployment, Intel raised capital spending after AI data-centre chip demand lifted revenue, and NVIDIA-backed projects in Korea outlined further large AI-factory expansion. This is likely to reinforce the line that AI demand is not behaving like a short burst around one model generation. The more immediate consequence for Microsoft, though, may be simpler: if compute stays expensive and abundant model supply keeps spreading across clouds, the durable advantage is likely to sit with whoever can package models into governed, costed and trusted operating systems for work.

Our forecast

Microsoft is likely to keep widening model choice inside its own enterprise surfaces while using in-house models first where they lower serving costs in Microsoft-run workloads.

due November 30, 2026 · what we hold ourselves to

By 2026-11-30, Microsoft must both (1) add at least one additional non-OpenAI third-party model after Claude Opus 5 to either Microsoft 365 Copilot or Microsoft Foundry, and (2) announce at least one further migration of a named Microsoft product or workload from an external model to a named MAI model.

Evidence
  • 2026-07-23 Microsoft launched its in-house generative AI models MAI-Image-2.5-Pro and MAI-Voice-2-Flash on Microsoft Foundry and began migrating workloads such as Bing Image Creator, PowerPoint image features, OneDrive image editing, Dynamics 365 Contact Center and Azure Voice Live from OpenAI models to these MAI models, reporting GPU cost reductions of up to 84% versus GPT-Image-2. techcommunity.microsoft.com
  • 2026-07-24 Microsoft added Anthropic’s Claude Opus 5 as a selectable model in Microsoft 365 Copilot across core apps, with rollout via the Copilot model selector depending on region and tenant configuration. techcommunity.microsoft.com
  • 2026-07-24 Microsoft made Anthropic’s Claude Opus 5 model available in Microsoft Foundry for building long‑running enterprise workflows and agentic applications, integrated with Foundry’s evaluation, security, governance and deployment tools. techcommunity.microsoft.com
  • 2026-07-22 Manulife signed a renewed and expanded five‑year partnership with Microsoft under which it will adopt the Microsoft 365 E7 Frontier Suite, deploy Microsoft Agent 365 for AI agent governance, and expand Microsoft 365 Copilot to more than 30,000 employees using Azure and Microsoft Foundry. news.microsoft.com
  • 2026-07-24 Microsoft released a CosmosMemoryContextProvider that uses Azure Cosmos DB as a native memory store for the Microsoft Agent Framework, providing durable cross‑session agent memory with automatic summarization and profiling, currently in Python preview. devblogs.microsoft.com
  • 2026-07-27 At a security event in San Francisco, Microsoft launched its MAI-Cyber-1-Flash cybersecurity AI model and the Project Perception agentic security platform, integrating the model into its MDASH multi-agent harness and claiming around 96% performance on the CyberGym benchmark at about half the cost of prior configurations. techcrunch.com
  • 2026-07-24 SAP decided to significantly restrict its spending on AI tokens after company‑wide AI usage drove token costs sharply higher, introducing a strict three‑stage controlling system to manage and limit AI token consumption. golem.de
  • 2026-07-28 Alibaba Cloud described an optimization to Qoder’s Quest 1.0 autonomous coding agent that reduces unnecessary MCP tool definitions in prompts, cutting context‑related token usage by over 10% in internal experiments while aiming to preserve agent capabilities. alibabacloud.com
  • 2026-07-22 OpenAI launched Presence, a managed enterprise platform for building, deploying, and operating governed AI agents for high‑volume, high‑stakes workflows, currently available only through limited managed deployments. help.openai.com
  • 2026-07-28 AWS announced that Amazon Bedrock AgentCore’s AgentCore Gateway now supports the Model Context Protocol (MCP) 2026‑07‑28 specification, introducing a stateless design, governed extensions, stronger authorization aligned with enterprise OAuth 2.0/OpenID Connect, and lifecycle guarantees configurable via UpdateGateway. aws.amazon.com
  • 2026-07-23 uniVersa insurance group reported a data protection incident in which an AI web crawler linked to OpenAI accessed a server during an IT migration and reached personal customer data including names, addresses, insurance numbers, tariff information and some bank details, after which the server was shut down, forensic experts were engaged and the case was reported to the Bavarian data protection authority. heise.de
  • 2026-07-27 404 Media reported that public share links from Anthropic’s Claude chatbot and its Artifacts feature were being indexed by Google, making users’ shared conversations and creations—including some sensitive health and company information—searchable on the open web. 404media.co
  • 2026-07-23 Codeberg e.V. updated its terms of use to prohibit using its services and hosted code as training data for generative AI or large language models and to allow exclusion of repositories that are largely autogenerated by AI agents or consume excessive resources, citing heavy load from AI crawlers and the energy and hardware demands of LLMs. heise.de
  • 2026-07-27 The EU’s AI Omnibus regulation entered into force on 27 July 2026, extending timelines and simplifying compliance for high-risk AI systems and SMEs under the AI Act, strengthening the AI Office’s enforcement powers, banning nudification apps that generate non-consensual sexual or intimate content and child sexual abuse material, and permitting certain sensitive-data processing to detect and correct bias in AI systems. digital-strategy.ec.europa.eu
  • 2026-07-27 NIST launched the Artificial Intelligence Technology Evaluation (AITE) program to provide a sequestered testbed for assessing AI model performance on blind data, initially focusing on image-analysis tasks using large vision-language models in domains such as quantum science, genomics and public safety. nist.gov
  • 2026-07-22 The Reserve Bank of India released for consultation a draft "Guidance on Regulatory Principles for Model Risk Management, 2026" setting governance and risk management requirements for traditional and AI/ML models used by regulated financial entities, including AI‑specific provisions on explainability, bias testing, hallucination controls, drift monitoring, human oversight, cybersecurity, customer disclosure and accountability for third‑party models. bloomberg.com
  • 2026-07-24 Google Cloud released Open Knowledge Format v0.2, adding optional frontmatter fields that provide provenance, trust, freshness, lifecycle and attestation signals so AI agents can assess and distinguish verified from unverified knowledge artifacts before consuming them. cloud.google.com
  • 2026-07-22 Alphabet reported Q2 2026 results showing 24% year‑over‑year revenue growth to $119.8 billion and 82% growth in Google Cloud revenue to $24.8 billion, attributing the cloud acceleration to demand for AI infrastructure and enterprise AI solutions. sec.gov
  • 2026-07-23 AMD launched its Helios AI rack-scale system and Instinct MI400 Series GPUs, alongside 6th Gen EPYC CPUs and Pensando networking, as a full-stack AI compute portfolio at its Advancing AI 2026 event. ir.amd.com
  • 2026-07-22 AMD and Anthropic announced a strategic partnership under which Anthropic plans to deploy up to 2 gigawatts of AMD Instinct MI450 Series GPUs in AMD Helios rack-scale solutions starting in 2027, while the companies collaborate on optimizing Claude workloads for AMD GPUs and AMD broadly adopts Claude internally. newsroom.amd.com
  • 2026-07-24 Intel reported second-quarter 2026 revenue of $16.1 billion, up 25% year on year and its fastest growth in about 15 years, driven largely by demand for chips used in AI data centres, and increased its capital spending plans by $2 billion. cn.ft.com
  • 2026-07-25 NAVER, NVIDIA and Brookfield announced plans to expand NAVER’s NVIDIA DSX-based AI factory at the GAK Sejong data centre in Korea from 55 megawatts to 200 megawatts by 2028 as part of gigawatt-scale multi-tenant AI cloud infrastructure investments, with NVIDIA to invest about $1 billion in NAVER and Brookfield to arrange up to $9 billion in project financing. investor.nvidia.com
  • 2026-07-25 SK Group and NVIDIA announced an expanded strategic partnership worth more than $500 billion that includes SK Telecom’s plan to build a 2-gigawatt NVIDIA Vera Rubin DSX AI factory in Korea and a long-term agreement for SK hynix and NVIDIA to co-develop and secure supply of future AI memory, including HBM. investor.nvidia.com