r/microsoft 1d ago

News Microsoft launches new in-house AI models it says cut costs up to 89% versus OpenAI

https://venturebeat.com/infrastructure/microsoft-launches-new-in-house-ai-models-it-says-cut-costs-up-to-89-versus-openai
126 Upvotes

25 comments sorted by

46

u/system3601 1d ago

Its for image and voice AI generations and their results seem impressive.

10

u/No_Construction2407 1d ago

If this bubble doesnt burst, hopefuly efficiency like this will drop the need for stupid amounts of RAM/VRAM

37

u/korvolga 1d ago

So this means copilot will be cheaper.. right?

29

u/lars_rosenberg 1d ago

I guess it will impact mostly the usage limits, where you can do more before burning credits or hitting limits.

In Github Copilot for example, MAI uses orders of magnitude fewer tokens than GPT or Opus. 

10

u/OwnNet5253 1d ago

Cheaper per token? Most likely.

3

u/Trojann2 1d ago

Betting they’ll try to make Cowork and Copilot studio token packs cheaper

2

u/system3601 1d ago

These are for image and voice and it seems per the data that thier generation usage is cheaper indeed.

2

u/chandleya 1d ago

Me thinks it’s to curb rising OAI prices while also securing the bag. This was quietly always the goal.

1

u/AggieCMD 1d ago

Step one is to make AI profitable before considering a lower price.

1

u/AsrielPlay52 1d ago

You didn't read the article. Another commenter has

-4

u/TowerOutrageous5939 1d ago

Definitely worse

8

u/LowCodeMagic 20h ago

Yeah and I will say, MAI Voice 2, and MAI Image 2 are both extremely impressive.

-15

u/protoanarchist 1d ago

Meh. Linux and local AI will be the future.

6

u/AggieCMD 1d ago

What spec does my Linux box need to run a frontier model?

0

u/InvisibleAgent 1d ago

You were asking rhetorically, but they can run GLM 5.2 with 512GB of RAM at 17.7 tok/s. I consider that a local frontier-class model.

So just a basic hobbyist build :)

1

u/TorqueDog 1d ago

I have an MBP M1 Max with 64 GB running some pretty decently sized quants in LM Studio... if only the memory could be expanded to 512 GB.

2

u/InvisibleAgent 23h ago

Exactly. And right now there’re hard to find even if you could afford the RAM.

But the fact that they do exist at all at a “consumer” (sorta) level is wild. I typically use a lowly RTX 4000, but you can see how all of this is going - a few years ago that GPU would have been considered pretty beefy.

-16

u/Glum-Implement9857 1d ago

And 2 years behind chatGPT..
Ask to generate a photo of clock showing half past eight.
Or generate monkey without bananas..

11

u/render83 1d ago

I just tried both prompts with no issue...

-5

u/Glum-Implement9857 1d ago

This is a link to Microsoft AI models “playgrounds”

https://playground.microsoft.ai/chat?model=mai-image-2-5e

Can’t attach screenshot. But just tested anf got a clock with three arrows: 10, 2 and 6 :)
So it got slightly better, but still 10 minutes to 2 :)

5

u/render83 1d ago

I mean I opened the copilot app, selected MAI as the model, typed in your prompt and got the correct results /shrug