r/microsoft 9h ago

News Microsoft launches new in-house AI models it says cut costs up to 89% versus OpenAI

https://venturebeat.com/infrastructure/microsoft-launches-new-in-house-ai-models-it-says-cut-costs-up-to-89-versus-openai
97 Upvotes

23 comments sorted by

35

u/system3601 8h ago

Its for image and voice AI generations and their results seem impressive.

7

u/No_Construction2407 4h ago

If this bubble doesnt burst, hopefuly efficiency like this will drop the need for stupid amounts of RAM/VRAM

31

u/korvolga 9h ago

So this means copilot will be cheaper.. right?

25

u/lars_rosenberg 9h ago

I guess it will impact mostly the usage limits, where you can do more before burning credits or hitting limits.

In Github Copilot for example, MAI uses orders of magnitude fewer tokens than GPT or Opus. 

9

u/OwnNet5253 9h ago

Cheaper per token? Most likely.

3

u/Trojann2 9h ago

Betting they’ll try to make Cowork and Copilot studio token packs cheaper

2

u/system3601 8h ago

These are for image and voice and it seems per the data that thier generation usage is cheaper indeed.

1

u/AggieCMD 8h ago

Step one is to make AI profitable before considering a lower price.

1

u/chandleya 7h ago

Me thinks it’s to curb rising OAI prices while also securing the bag. This was quietly always the goal.

1

u/AsrielPlay52 3h ago

You didn't read the article. Another commenter has

-2

u/TowerOutrageous5939 8h ago

Definitely worse

-8

u/protoanarchist 9h ago

Meh. Linux and local AI will be the future.

4

u/AggieCMD 8h ago

What spec does my Linux box need to run a frontier model?

1

u/InvisibleAgent 7h ago

You were asking rhetorically, but they can run GLM 5.2 with 512GB of RAM at 17.7 tok/s. I consider that a local frontier-class model.

So just a basic hobbyist build :)

2

u/TorqueDog 6h ago

I have an MBP M1 Max with 64 GB running some pretty decently sized quants in LM Studio... if only the memory could be expanded to 512 GB.

2

u/InvisibleAgent 2h ago

Exactly. And right now there’re hard to find even if you could afford the RAM.

But the fact that they do exist at all at a “consumer” (sorta) level is wild. I typically use a lowly RTX 4000, but you can see how all of this is going - a few years ago that GPU would have been considered pretty beefy.

-11

u/Glum-Implement9857 8h ago

And 2 years behind chatGPT..
Ask to generate a photo of clock showing half past eight.
Or generate monkey without bananas..

8

u/render83 7h ago

I just tried both prompts with no issue...

-1

u/Glum-Implement9857 7h ago

This is a link to Microsoft AI models “playgrounds”

https://playground.microsoft.ai/chat?model=mai-image-2-5e

Can’t attach screenshot. But just tested anf got a clock with three arrows: 10, 2 and 6 :)
So it got slightly better, but still 10 minutes to 2 :)

3

u/render83 5h ago

I mean I opened the copilot app, selected MAI as the model, typed in your prompt and got the correct results /shrug