๊ฐ€์ž… ํ›„ ์ดˆ๋Œ€ ๋งํฌ๋ฅผ ๊ณต์œ ํ•˜๋ฉด ๋™์˜์ƒ ์žฌ์ƒ ๋ฐ ์ดˆ๋Œ€ ๋ณด์ƒ์„ ๋ฐ›์„ ์ˆ˜ ์žˆ์Šต๋‹ˆ๋‹ค.

Google Gemma
@googlegemma
The official home of Google's Gemma. Lightweight, state-of-the-art open models by Google DeepMind, built on Gemini tech. What will you build? ๐Ÿš€๐Ÿ’ป
๊ฐ€์ž… April 2025
0 ํŒ”๋กœ์ž‰ ์ค‘    101.1K ํŒฌ
"DiffusionGemma as Jev" showcases the power of non-autoregressive architectures. While Jev demonstrates the value of rapid decision models, running DiffusionGemma in this paradigm leverages canvas diffusion to evaluate structured choices in a single parallel pass: โšก ๏ธMassive Parallelism: Denoises across an open canvas in a single step instead of sequential autoregressive token generation (~0.2s on a DGX spark). ๐Ÿง  Full Bidirectional Attention: Allows every option to attend to the full context concurrently, yielding well-calibrated decision distributions. ๐Ÿ‘๏ธ Multimodal Grounding: Inherits Gemma 4's spatial vision capabilities for complex visual and text decisions. Read more about this approach here:
๋” ๋ณด๊ธฐ
I ran some real, live evals on Jev vs DiffusionGemma-as-Jev (my patch for vLLM!) DiffusionGemma comes out as the winner, I think. Headlines: Is Jev faster than DiffusionGemma? No โŒ (API vs DGX Spark) Is Jev smarter than DiffusionGemma? No โŒ (they're roughly tied!)
๋” ๋ณด๊ธฐ