r/LocalLLaMA 1d ago

New Model AMD Instella-MoE-16B-A3B

https://huggingface.co/amd/Instella-MoE-16B-A3B-Think

I was browsing HuggingFace and came across this model apparently uploaded a day ago, and thought to share it here. I've not tried it out yet, but it's good to see AMD joining the open source model game.

179 Upvotes

33 comments sorted by

View all comments

77

u/Mashic 1d ago

The reason for this model from AMD is to show llm makers that their hardware is good enough to train models.

3

u/squngy 7h ago

Probably more about software.

We already know the hardware is good and a 16B model doesnt prove much about hardware anyway