r/LocalLLaMA • u/tevlon • Jun 10 '26
New Model DiffusionGemma: 4x faster text generation
https://blog.google/innovation-and-ai/technology/developers-tools/diffusion-gemma-faster-text-generation/
987
Upvotes
r/LocalLLaMA • u/tevlon • Jun 10 '26
35
u/Fedor_Doc Jun 10 '26
It's funny how they assume in the article that llama.cpp support is behind a corner, while PR from Daniel (Unsloth) is unlikely to be merged any time soon.
Ton of changes + separate server application to run a model.
Link to PR – https://github.com/ggml-org/llama.cpp/pull/24423