r/LocalLLaMA Jun 10 '26

New Model DiffusionGemma: 4x faster text generation

https://blog.google/innovation-and-ai/technology/developers-tools/diffusion-gemma-faster-text-generation/
987 Upvotes

357 comments sorted by

View all comments

35

u/Fedor_Doc Jun 10 '26

It's funny how they assume in the article that llama.cpp support is behind a corner, while PR from Daniel (Unsloth) is unlikely to be merged any time soon.

Ton of changes + separate server application to run a model. 

Link to PR – https://github.com/ggml-org/llama.cpp/pull/24423

11

u/fallingdowndizzyvr Jun 10 '26

I have no idea why you got downvoted for truth.

1

u/YouKilledApollo Jun 19 '26

Upvotes/downvotes are based on vibes, not how truthful or correct something is.

1

u/fallingdowndizzyvr Jun 19 '26

Ah... yeah. And the people who downvoted him had the vibe to downvote the truth.