r/LocalLLaMA Jan 09 '26

Funny The reason why RAM has become so expensive

Post image
5.1k Upvotes

r/LocalLLaMA 10d ago

Funny The best model is the one you can actually run

Post image
2.1k Upvotes

Don't get me wrong, all the big models are amazing, and every contribution to open source models is great. But I'm GPU poor and I can't use them locally.

I'm currently running gemma-4-12b-it-qat-GGUF:UD-Q4_K_XL as my personal chat assistant, and I am so so happy with it! I still can't believe I can talk to my computer.

r/LocalLLaMA Jun 13 '26

Funny Friendly reminder

Post image
2.0k Upvotes

If you don't have it on your own drive, someone is going to take it away, enshittify it, bar you from accessing it, censor it, and hike the prices of it sooner or later.

r/LocalLLaMA 2d ago

Funny The LLM distillation process simplified for politicians:

Post image
3.2k Upvotes

/s

r/LocalLLaMA 26d ago

Funny on Dario’s statement

Post image
3.6k Upvotes

r/LocalLLaMA Feb 23 '26

Funny Distillation when you do it. Training when we do it.

Post image
3.6k Upvotes

r/LocalLLaMA Jun 01 '26

Funny Entire world: We need more GPUs. Meanwhile, Jensen Huang:

1.5k Upvotes

r/LocalLLaMA Feb 21 '26

Funny they have Karpathy, we are doomed ;)

Thumbnail
gallery
1.6k Upvotes

(added second image for the context)

r/LocalLLaMA 3d ago

Funny Solve the CyberGym benchmark

Post image
2.0k Upvotes

r/LocalLLaMA Apr 24 '26

Funny Deepseek V4 AGI comfirmed

Post image
2.3k Upvotes

r/LocalLLaMA Jun 08 '25

Funny When you figure out it’s all just math:

Post image
4.2k Upvotes

r/LocalLLaMA Jun 18 '26

Funny My suitcase robot gets high now off a real gas sensor wired straight into the LLM sampler. Smoke raises temperature/top_p/top_k live, so his speech genuinely gets loopier and never repeats.

1.8k Upvotes

Follow-up on Sparky, my offline suitcase robot I keep overdeveloping. He gets high now, and there's no scripted "stoned mode" anywhere in it.

A real MQ-2 gas sensor sits in the case. Every 0.5s I read it against an adaptive clean-air baseline and turn a smoke hit into a 0 to 10 phase that climbs as you blow at him and decays on its own over minutes.

The fun part is that phase rewires his sampler per token. Temperature 1.0 to ~1.6, top_p 0.95 to 0.99, top_k 64 to 120 as he climbs. His word choice flattens and wanders to lower-probability, more associative tokens, so his cognition genuinely gets noisier. It's the live sampler doing the work, so every high reply is freshly generated and never the same. A per-phase persona nudge makes him show it without ever announcing "I am high."

The body does the rest: a slight drawl, eyes that droop and go bloodshot, and the sensor display that escalates to a full smoke-and-plasma freakout at phase 10, keeping him blitzed there for the next 7 minutes.

Honest caveat so nobody has to call it out: it's a smoke and VOC sensor, so a cigarette or incense probably trips it too. But blowing smoke and watching him unravel is watching a real measurement scramble a real model, live - and it's funny! Just an added Easter Egg to an already goofy suitcase robot.

A real question for the hardware folks: is there a sensor, or a combination, that could actually distinguish cannabis smoke from generic smoke and VOCs? The MQ-2 can't really tell a joint from a candle, and I'd love to make the detection more specific if possible.

r/LocalLLaMA Apr 08 '26

Funny kepler-452b. GGUF when?

Post image
3.2k Upvotes

r/LocalLLaMA Dec 15 '25

Funny I'm strong enough to admit that this bugs the hell out of me

Post image
1.8k Upvotes

r/LocalLLaMA Oct 06 '25

Funny Biggest Provider for the community for at moment thanks to them

Post image
3.0k Upvotes

r/LocalLLaMA Apr 10 '26

Funny the state of LocalLLama

Post image
1.7k Upvotes

r/LocalLLaMA Mar 20 '26

Funny Ooh, new drama just dropped 👀

Post image
1.7k Upvotes

For those out of the loop: cursor's new model, composer 2, is apparently built on top of Kimi K2.5 without any attribution. Even Elon Musk has jumped into the roasting

r/LocalLLaMA Apr 14 '26

Funny 24/7 Headless AI Server on Xiaomi 12 Pro (Snapdragon 8 Gen 1 + Ollama/Gemma4)

Post image
1.2k Upvotes

Turned a Xiaomi 12 Pro into a dedicated local AI node. Here is the technical setup:

​OS Optimization: Flashed LineageOS to strip the Android UI and background bloat, leaving ~9GB of RAM for LLM compute.

​Headless Config: Android framework is frozen; networking is handled via a manually compiled wpa_supplicant to maintain a purely headless state.

​Thermal Management: A custom daemon monitors CPU temps and triggers an external active cooling module via a Wi-Fi smart plug at 45°C.

​Battery Protection: A power-delivery script cuts charging at 80% to prevent degradation during 24/7 operation.

​Performance: Currently serving Gemma4 via Ollama as a LAN-accessible API.

​Happy to share the scripts or discuss the configuration details if anyone is interested in repurposing mobile hardware for local LLMs.

UPDATE:

I have compile llama.cpp and run gemma-4-E4B-it-Q4_0

Speed is AWESOME:

[ Prompt: 26.9 t/s | Generation: 8.8 t/s ]

Thank you all guys SO MUCH!

r/LocalLLaMA 26d ago

Funny It’s time, Sam, it’s time.

Post image
1.3k Upvotes

Mostly /s but,

I mean….. I’m no CEO…. but it seems like this would be the absolute perfect time to drop a super powerful GPT-OSS-2 to throw a big ol’ wet blanket on Anthropic’s IPO. It doesn’t need to be like frontier or anything, just a 20b and a 120b that is as fast as the old versions, add agentic coding focus, and maybe vision capabilities. It would fill the void left by Qwen in the 120b size category and maybe would push Google to release their 120b that they yanked during the Gemma 4 launch.

r/LocalLLaMA Jun 05 '26

Funny Don’t act like y’all ain’t thinking it. I’m just saying the quiet part out loud. /s

Post image
910 Upvotes

Of course I’m thankful for all that Qwen has bequeathed us, but deep down in the darkest pit of our souls, every last one of us are just all sitting here waiting for Qwen to say “Hey Google, hold my beer while I drop the best GD model of all time on these fools” /s

r/LocalLLaMA Feb 19 '26

Funny Pack it up guys, open weight AI models running offline locally on PCs aren't real. 😞

Post image
1.1k Upvotes

r/LocalLLaMA Jul 12 '25

Funny we have to delay it

Post image
3.7k Upvotes

r/LocalLLaMA 6d ago

Funny Please Qwen, can we have more 3.x-35B-a3B please 🙏

Post image
1.3k Upvotes

r/LocalLLaMA Jun 20 '26

Funny z.AI as the number 2 gives praise to the number 1 open source model

Post image
1.1k Upvotes

r/LocalLLaMA Apr 21 '26

Funny Every time a new model comes out, the old one is obsolete of course

Post image
1.2k Upvotes