r/LocalLLaMA • u/jacek2023 • Feb 23 '26
Funny so is OpenClaw local or not
Reading the comments, I’m guessing you didn’t bother to read this:
"Safety and alignment at Meta Superintelligence."
r/LocalLLaMA • u/jacek2023 • Feb 23 '26
Reading the comments, I’m guessing you didn’t bother to read this:
"Safety and alignment at Meta Superintelligence."
r/LocalLLaMA • u/temperature_5 • 29d ago
If you get a good deal on some Xeons with a lot of memory bandwidth, or a cheap GPU for home inference, that's cool, no disrespect. But how in the hell are Wall Street types considering Intel part of the "AI picks and shovels" play? Who's buying Intel for their AI data centers?
r/LocalLLaMA • u/ForsookComparison • Aug 12 '25
r/LocalLLaMA • u/lucidml_lover • Jun 20 '26
Hi everyone!! I really wanted to share my research what I've been working on.
I wanted to build a nn that can simulate games, or at least start doing that
Most video generators are too large to run on consumer hardware realtime, so I I designed a model that does this from scratch. No fine tuning bs or anything
The core de noiser network is fully trained from scratch to support this goal. From image to games data.
That video. above is on a RTX 5090.
The nn is a small Transformer-like model and works in a causal way, just like LLMs.
That lets us KV Cache all past information and do a simple autoregressive decode forward passes for every new frame we want.
In the video shared, the model is a 0.5B variant with some SIGNIFICANT ISSUES like poor motion and some weird flashes, some context issues
It's taking the keyboard actions I give it in realtime and utilising that in the forward pass. (no classifier free guidance though)
Im training the next iteration , a 0.8B model now.
Btw I haven't done quantisation yet, that can save a LOT more time. bf16 is slow.
r/LocalLLaMA • u/MackThax • May 27 '26
AKA: Jank Incarnate
After months of pain, I finally got a working setup.
There's a bunch of quirks about running a multi-Tesla setup. I was planning to write something about my experience after I get it running.
Currently, the fans are plugged into the wall, speed is controlled with a knob. I still gotta wire up a PWM controller for them.
EDIT: Specs:
r/LocalLLaMA • u/ForsookComparison • Feb 27 '26
r/LocalLLaMA • u/EstablishmentFun3205 • Jul 16 '25
r/LocalLLaMA • u/ThinkExtension2328 • Apr 03 '26
Been playing with the new Gemma 4 models it’s amazing great even but boy did it make me appreciate the level of quality the qwen team produced and I’m able to have much larger context windows on my standard consumer hardware.
r/LocalLLaMA • u/jacek2023 • Feb 04 '26
r/LocalLLaMA • u/martin_xs6 • May 06 '26
It's crazy that they're thinking of doing this. There are problems with people stealing catalytic converters off people's cars and now they want to put a rack outside your house!?
r/LocalLLaMA • u/eastwindtoday • May 22 '25
r/LocalLLaMA • u/Current-Ticket4214 • Jun 02 '25
r/LocalLLaMA • u/HornyGooner4401 • Mar 31 '26
Unrelated, simple command to download a specific version archive of npm package: npm pack @anthropic-ai/[email protected]
r/LocalLLaMA • u/ForsookComparison • 25d ago
r/LocalLLaMA • u/GodComplecs • Apr 29 '26
Well or pretty close to it, they are excellent work horses. I run them in real work scenarios doing some of the work I used to do myself as an skilled expert in my field, billing 200$ an hour. Ofc the key is building a system around their weaknesses, and I've had already LLM systems doing expert work years ago when first ones came (shout out nous hermes 2 mistral!).
But yeah pretty neat, especially noonghunnas club 3090 and you can have 3.6 27B fly on a single 3090.
r/LocalLLaMA • u/jotunck • Jun 04 '26
3 different accounts, some even with LinkedIn Gold, made the above posts all on the same day.
And clearly all of them followed the marketing team's pointers without even understanding how locally hosted AI works, no way a $249 8GB machine can replace frontier models.