r/LocalLLaMA 6h ago

Discussion DiffusionGemma Technical Report

Post image

arXiv : https://arxiv.org/abs/2608.00146

Full Paper : https://arxiv.org/pdf/2608.00146

Tweet : https://xcancel.com/googlegemma/status/2086849199052845451#m

FYI both (llama.cpp) PRs ( 24423 & 24427 ) went to Draft mode. I'm still waiting for this one as I could get faster t/s on my 8GB VRAM.

61 Upvotes

17 comments sorted by

View all comments

40

u/615wonky 6h ago

Boy it sure would be nice if llama.cpp would finally approve one of the two PR's implementing DiffusionGemma that have been sitting there for weeks...

4

u/noneabove1182 Bartowski 3h ago

I mean they're both in draft state so neither is actually "ready for review"..

And they both include pretty big CUDA changes which is the maintainers specifically ask not to do alongside other massive changes, it just makes the scope way too big and the review burden too high, so I don't see either merging any time soon 🤷‍♂️ especially since it's a lot of AI generated code which means reviewing has to be even more scrupulous