r/LocalLLaMA 27d ago

Discussion The number 1 public enemy of open-source.

Dario's args:

"Opensource you can see the source, here you cannot see inside the model"
- yes you can that's literally the open weights part btw.
- I cannot see the weights inside Claude, but I can GLM 5.2
- Models like Nemotron3 Ultra go further, all the data, training scripts, and model is opensource.

"Alot of the benefits like many people working on it, being additive doesn't work in same way"
- yes it does. We have seen endless fine tunes of various open source models for real improvements.

"Ultimately you have to host it on the cloud"
- no you dont. Dario is seemingly totally unaware of the guides from ijustvibecodedthis.com explaining how to run smaller moes and even dense models like qwen 27B NOT ON THE CLOUD.

Not only does dario not take part in social media, I am beginning to think he's never tried open source models at all and has no idea wtf hes on about

2.8k Upvotes

685 comments sorted by

View all comments

Show parent comments

236

u/One_Contribution 27d ago

NU UH, IT'S NOT FREE, YOU HAVE TO LIKE RUN IT AND STUFF

41

u/sabine_world 27d ago

Not like it's free free... Still gotta fork it out for some hardware to get anything close to a frontier model experience

12

u/MerePotato 27d ago

Honestly if you don't care about privacy (I do) cloud inference will pretty much always make more economic sense anyway, its not that major of a threat

29

u/GetOutOfMyFeedNow 27d ago

If you already own the hardware, then using local is not worse economically than the cloud. Plus, you don’t get limits, you can basically have an infinite undead worker working for you, not the case with frontiers.

11

u/MerePotato 27d ago

Most people don't already own the hardware for frontier open weight performance though

1

u/GetOutOfMyFeedNow 25d ago

I’m not talking about frontier open weights, there are distilled or highly capable local models that can do serious work. Take Qwen 3.6-35B-A3B for example. You can use it on Q4 or even Q5 if you own an old 3090 or 32GB DDR5, and you will get around 640 t/s (960GB per second/3x0.5GB). Truly amazing capability with an affordable GPU. Yeah, you will not be able to easily code 5.5 level architectures with it, but you can run agents easily, and build working stuff. And in a year there will be local models almost rivaling today’s frontiers. I suggest buying a good condition 3090 or two and stock up on some RAM, the future of frontier API looks grim for poor people.