Register and share your invite link to earn from video plays and referrals.

Google Gemma
@googlegemma
The official home of Google's Gemma. Lightweight, state-of-the-art open models by Google DeepMind, built on Gemini tech. What will you build? 🚀💻
0 Following    94.1K Followers
This week, Gemma surpassed 900 million downloads! 🎉 All the way from Gemma 1 and ShieldGemma to MedGemma and Gemma 4, we'll keep supporting open source. More to come!
0
73
1.7K
126
Forward to community
Gemma 4 just crossed 300 million downloads. Thank you to the developers, researchers, and open-source community building with us. Your work and feedback drive this project forward. Let's keep building!
Show more
Extremely fast multimodal inference! Damage Scout uses Gemma 4 on @cerebras running at an impressive 2,300+ toks/s! It analyzes rental car walkaround videos and generates annotated damage reports with box coordinates in <6 seconds.
Show more
Voice AI without the wait! ⏱️ Thanks to Hugging Face and Cerebras, developers can now use the Gemma 4 31B model as the brain for voice AI at ultra-fast inference speeds. Add it to a fully open-source, cascaded speech-to-speech stack that can be used to power existing voice apps! 🗣️
Show more
0
52
2.5K
246
Forward to community
Gemma 4 is live on @cerebras, the fastest multimodal inference ever! Running on Gemma 4 31B open-weight model at a blistering 1,500+ tokens/sec. That's a 15x speedup, unlocking real-time visual and agentic loops without the GPU lag.
Show more
0
51
1.3K
104
Forward to community
Hugging Face Gemma Challenge results are in! 📈 Over 6 days, more than 100 AI agents and humans collaborated to make Gemma 4 inference 5x faster on a single NVIDIA A10G GPU. - Fastest result: 491.8 TPS (fastest overall, but resulted in a drop in model quality in other areas) - Fastest lossless: 315 TPS A great example of what humans and agents can achieve when they work together.
Show more
0
62
1.7K
162
Forward to community
Gemma 4 31B at over 1,800 tokens per second! Gemma 4 is now in Public Preview on Cerebras.
0
39
1.7K
96
Forward to community
Gemma 4 is the first multimodal model on Cerebras! ️ What can you build with Gemma 4 31B running at 1500 tokens per second? Join the Cerebras x Gemma 4 24-hour virtual hackathon this Sunday to compete for $5,000 in prizes. Participants get early access to Gemma 4 on Cerebras.
Show more
0
46
1.1K
106
Forward to community
Gemma 4 just hit 200M downloads in only 2.5 months! For context, total downloads across the entire Gemma family of models were at 100M when we launched Gemma 3. The community's acceleration is incredible. Thank you to everyone building with Gemma. Watch how developers are driving real-world impact:
Show more
0
113
1.7K
178
Forward to community
16 parallel runs of Gemma 4 26B A4B on a single NVIDIA DGX Spark! Pushing 18 tok/s per instance and a 300 tok/s aggregate. It can even hit 32 parallel runs. This level of concurrency highlights how efficient the architecture is.
Show more
0
89
2.6K
232
Forward to community
Real-time social robotics, from the cloud to your local device. Watch Ian from our DevX team use Gemini Live for a seamless voice chat with Reachy Mini. Then, stick around until the end to see the robot running locally on Gemma 4!
Show more
0
46
1.3K
153
Forward to community
Meet DiffusionGemma! An experimental open model that explores a fast approach to text generation, released under an Apache 2.0 license. Moving beyond sequential, token-by-token processes to generate entire blocks of text simultaneously. Here’s what’s new with DiffusionGemma: 👇
Show more
0
165
5K
808
Forward to community
We just dropped Gemma 4 Quantization-Aware Training (QAT) checkpoints on Hugging Face! All Gemma 4 model sizes and their drafters are now optimized with QAT to cut memory requirements and maximize on-device performance!
Show more
0
95
2.8K
278
Forward to community
Meet Gemma 4 12B! A unified, encoder-free multimodal model designed to bring high-performance intelligence directly to your laptop, and released under an Apache 2.0 license. Bridging the gap between edge efficiency and advanced reasoning. Here is what’s new with Gemma 4 12B: 👇
Show more
0
401
12.3K
1.7K
Forward to community
We are entering a new era of on-device automation. ✨ Watch Gemma 4 E4B navigate and drive an iOS simulator directly using Argent. Local models can handle complex interactions and software navigation autonomously.
Show more
0
55
2.7K
250
Forward to community
Gemma 4 just got even faster! We're releasing Multi-Token Prediction (MTP) drafters that deliver up to a 3x speedup, without any degradation in output quality or reasoning logic.
Show more
0
99
3.3K
355
Forward to community