Sources: Microsoft is training MAI-1, a new, in-house AI model with ~500B parameters, large enough to compete with top models from Google, Anthropic, and OpenAI
The InformationAaron Holmes
Context & Ripple Effects
Microsoft’s reported MAI-1 training marks an early move from relying solely on external frontier models toward owning a comparable model layer. The subsequent record shows that this was not a one-off experiment: Microsoft later completed a family of MAI models and tested them in Copilot.
The strategy broadened beyond a single general-purpose model into speech, image, and reasoning systems, including in-house MAI models for transcription, voice, and image generation. That makes MAI-1 consequential as a foundation for Microsoft’s control over the models deployed across its products.
First-order effects
Microsoft gains a potential in-house model candidate at the scale of leading offerings from Google, Anthropic, and OpenAI, rather than being limited to external model suppliers.
OpenAI’s position inside Microsoft products becomes less exclusive in strategic terms: MAI-1 creates an internal alternative even before any product substitution occurs.
Second-order effects
Microsoft can use an internal model option to benchmark external providers on capability, cost, and product fit; later reporting that it began replacing some OpenAI and Anthropic models in Excel and Outlook illustrates the commercial leverage this architecture can create.
Google, Anthropic, and OpenAI face a customer that is also building a competing supplier, increasing the importance of maintaining differentiated model performance and terms for large platform distributors.
Third-order effects
If major AI distributors continue developing their own frontier models, the market can shift from a small set of specialist model vendors toward vertically integrated platforms that both train models and control end-user distribution.
The durable constraint becomes less simply access to a frontier model and more the ability to operate a portfolio of proprietary and third-party models—an industrialization path that may preserve external suppliers where they remain distinctly better.
The trend: MAI-1 is an early data point in the vertical integration of AI, as large software platforms build proprietary models to reduce supplier dependence while retaining flexibility to use outside systems.
Assuming this is correct (no idea) it is interesting that both Meta and MS have decided on ~400B - 500B param (assuming dense) models. Also interesting that is also roughly the same amount of activated params as the rumored GPT4 architecture. doesn't mean much by itself, but it..…
Today's big scoop from @theinformation is that Microsoft is training a giant model to compete with OpenAI. This fits into a larger pattern, going back over a year. Nadella clearly doesn't want to be dependent on OpenAI. And because OpenAI doesn't appear to have a lot of unique...
Microsoft is now working on its own Model (LLM) MAI-500B. My Thoughts 👇🏻 No surprise that Microsoft, despite its deep connection to OpenAI, is going to develop and build its own LLMs. What I don't see is this as an “attack” of any sort on OpenAI, but rather a diversification of…
As I predicted, Microsoft is training its own LLM It's called MAI-1, a 500B param model, and may be previewed at their Build conference. When this model becomes available, it will only be natural for MSFT to push this instead of the GPT line. As predicted, OAI and MSFT are... [im…
Much industry talk today about @Microsoft training it's own #LLM. Known as #MAI-1, it has 500 billion parameters. The big Q is how @OpenAI will be positioned in the mix. My take: Microsoft has too few variety of models + this will diversify. But will alter OpenAI partnership.
Scoop: Microsoft is training its own large language model, internally labaled MAI-1, with Mustafa Suleyman leading the effort. The model is around 500B parameters and could compete directly with LLMs from Google, OpenAI, etc. Details: https://www.theinformation.com/ ...
@bindureddy While a smart model is important... the ability to build an agent framework around it is the only way to provide value (especially at the enterprise scale) MSFT has the absolute worst framework to do this... and their bloated ecosystem will never be as streamlined and…
There have been indications for some time that Microsoft was intending to take their own path, and were planning to train a bespoke model to compete with GPT-4. We now have a name. ‘MAI-1’ As well as a size: 500b.
AI Agenda scoop from @aaronpholmes: Microsoft is training its own LLM, led by Mustafa Suleyman. The model is ~500Bn parameters, showing how MSFT is dual tracking its AI efforts, building both small, cheap models and larger ones that compete with OpenAI. https://www.theinformation…
I mentioned this not long after the Open AI relationship was clear last year but every company wanting to be an AI platform can't afford to outsource their primary model. This is the easiest prediction of vertical integration there is at the moment.
In a LinkedIn post just now, Microsoft CTO Kevin Scott confirms my report this morning that MAI is Microsoft's new AI model in the works (while tamping down hype): [image]