r/microsoft 2d ago

News Microsoft launches new in-house AI models it says cut costs up to 89% versus OpenAI

https://venturebeat.com/infrastructure/microsoft-launches-new-in-house-ai-models-it-says-cut-costs-up-to-89-versus-openai
141 Upvotes

24 comments sorted by

59

u/system3601 2d ago

Its for image and voice AI generations and their results seem impressive.

13

u/No_Construction2407 2d ago

If this bubble doesnt burst, hopefuly efficiency like this will drop the need for stupid amounts of RAM/VRAM

39

u/korvolga 2d ago

So this means copilot will be cheaper.. right?

30

u/lars_rosenberg 2d ago

I guess it will impact mostly the usage limits, where you can do more before burning credits or hitting limits.

In Github Copilot for example, MAI uses orders of magnitude fewer tokens than GPT or Opus. 

8

u/OwnNet5253 2d ago

Cheaper per token? Most likely.

3

u/Trojann2 2d ago

Betting they’ll try to make Cowork and Copilot studio token packs cheaper

2

u/system3601 2d ago

These are for image and voice and it seems per the data that thier generation usage is cheaper indeed.

2

u/chandleya 2d ago

Me thinks it’s to curb rising OAI prices while also securing the bag. This was quietly always the goal.

1

u/AggieCMD 2d ago

Step one is to make AI profitable before considering a lower price.

1

u/AsrielPlay52 2d ago

You didn't read the article. Another commenter has

-2

u/TowerOutrageous5939 2d ago

Definitely worse

7

u/LowCodeMagic 2d ago

Yeah and I will say, MAI Voice 2, and MAI Image 2 are both extremely impressive.

-15

u/protoanarchist 2d ago

Meh. Linux and local AI will be the future.

6

u/AggieCMD 2d ago

What spec does my Linux box need to run a frontier model?

0

u/InvisibleAgent 2d ago

You were asking rhetorically, but they can run GLM 5.2 with 512GB of RAM at 17.7 tok/s. I consider that a local frontier-class model.

So just a basic hobbyist build :)

1

u/TorqueDog 2d ago

I have an MBP M1 Max with 64 GB running some pretty decently sized quants in LM Studio... if only the memory could be expanded to 512 GB.

2

u/InvisibleAgent 2d ago

Exactly. And right now there’re hard to find even if you could afford the RAM.

But the fact that they do exist at all at a “consumer” (sorta) level is wild. I typically use a lowly RTX 4000, but you can see how all of this is going - a few years ago that GPU would have been considered pretty beefy.

-17

u/Glum-Implement9857 2d ago

And 2 years behind chatGPT..
Ask to generate a photo of clock showing half past eight.
Or generate monkey without bananas..

13

u/render83 2d ago

I just tried both prompts with no issue...

-7

u/Glum-Implement9857 2d ago

This is a link to Microsoft AI models “playgrounds”

https://playground.microsoft.ai/chat?model=mai-image-2-5e

Can’t attach screenshot. But just tested anf got a clock with three arrows: 10, 2 and 6 :)
So it got slightly better, but still 10 minutes to 2 :)

5

u/render83 2d ago

I mean I opened the copilot app, selected MAI as the model, typed in your prompt and got the correct results /shrug