Register and share your invite link to earn from video plays and referrals.

NVIDIA AI
@NVIDIAAI
Teaching your AI new tricks.
898 Following    343.5K Followers
Ten years ago at NVIDIA GTC 2016, our CEO @JensenHuang introduced DGX-1, our first AI supercomputer. Since then, NVIDIA DGX has transformed from a single system into a blueprint for AI factories. What hasn’t changed is the reason why we build DGX: to test the systems and work through the details of interconnects, cooling, and software so our partners and customers have a proven foundation to build on. Charlie Boyle, our VP of DGX Systems, looks back at how DGX got its start, what’s changed, and what comes next. Thanks to everyone who’s been on this journey with us. #DecadeOfDGX# Watch the full video:
Show more
Love this news. More doctors around the world will be able to use OpenEvidence for free. Congrats to @OpenEvidence and @AnthropicAI 💚
“Medicine is a field where you don't need to squint to see the benefits of AI. It's your mom, it's your dad, it's your sister, it's your brother. And it's the doctors who will care for them in their most vulnerable moments.” Yesterday, during the UN General Assembly, we announced a partnership with @AnthropicAI to make OpenEvidence available free of charge to physicians in low- and middle-income countries, placing life-saving medical AI into the hands of physicians worldwide, regardless of the economics.
Show more
🚀 New benchmark alert: SWE-Serve ( Can AI agents develop inference engine and make it serve real models? Our team at @nvidia built 53 tasks from real @sgl_project engineering work to find out. The chart shows why live-serving tests matter. 🧵
Show more
Our new Nemotron 3 Diarization model + @pollenrobotics Reachi Mini + DGX Spark Great work Andi @huggingface 🙌
Voice agents still don’t understand who’s speaking to them. That’s a huge gap compared with humans, hidden by all the “phone-call” demos. But that changes today! NVIDIA is open-sourcing Nemotron 3 Diarization: a model that can reliably track speakers in live conversations, under a commercial-friendly license! In my tests, the quality is really good with one-second speech chunks. So we can use it for voice agents! I tested it with Reachy Mini and speech-to-speech running on a DGX Spark. It’s super fun to see the robot notice a new voice, ask for a name, and remember it. The model has day-zero integration with Transformers! Kudos to the NVIDIA team for shipping useful tools for the whole community!
Show more
When several people talk at once, a transcript can get messy fast. Our new Nemotron 3 Diarization model tracks who spoke when, even when voices overlap. It handles up to eight speakers, has 100M parameters, and is now available on @huggingface 🤗
Show more
0
189
9.6K
856
Forward to community
Figuring out a data center’s carbon footprint means piecing together a lot of messy data, from parts lists and documents to supplier emails. Sluicebox built an AI supplier agent called Lucy to help. Using NVIDIA Nemotron 3 Ultra, Lucy collects and checks supplier data so engineers can understand the impact of their choices while they’re still designing. In Sluicebox’s tests, Nemotron improved accuracy, cut costs by 51–80% and delivered responses up to 2x faster. See how it works:
Show more
Ask the Experts: Evaluating Agent Skills | Nemotron Labs
vLLM integrated PyNvVideoCodec to offload video decoding from CPU to GPU's NVDEC. Result: 2x+ throughput at 8×H100, CPU bottleneck gone! This is huge for video captioning at scale (AV training, metadata), and ships with CUDA vLLM releases. 🔗
Show more
And the winners are .... Congratulations to our #NVIDIAGTC# Berlin Golden Ticket winners! Chosen by our panel of judges @asierarranz @choroukmalmoum @SteveNouri @mervenoyann @itsjohnnynunez and @googlecloud, these six winners will join us in Berlin next month: 🏆 Lucas Hudson 🏆 James Kane 🏆 Devin Nicholson 🏆 Aaron Nowak 🏆 Vladimir Shirokun 🏆 Stefan Trauth Thanks to everyone who participated and shared your projects - we loved seeing what this community is building with open models. See you at GTC!
Show more
What are you building with open models? Show us what you’re working on and you could be headed to #NVIDIAGTC# Berlin. 🎫 Your Golden Ticket includes: → Free conference pass → VIP seating for Jensen’s keynote → Exclusive NVIDIA merch → Access to special events Submissions are open Aug 18 – Sep 10, 2026. Details and how to enter:
Show more
Your LLM endpoint works. But how does it perform when traffic increases? NVIDIA Dynamo AIPerf helps you measure TTFT, ITL, latency and throughput at scale, then test with realistic traffic patterns you can reliably repeat. Read the blog:
Show more
Seattle Spark Hack Winners Livestream Spotlight: LiveKit & Memo
What happens when you give a world model 32 images of a place we know very well? World Labs put Atlas to the test on Voyager. Atlas brings text, images, video, and 3D into a shared spatial context, enabling one model to generate new views, reconstruct scenes, and simulate worlds. With Voyager, you can now see that in action in real time. Pretrained from scratch on NVIDIA Blackwell GPUs. Take a look around 👇
Show more
At Dumfries House in Scotland, our CEO @JensenHuang joined His Majesty The King and global leaders to discuss how AI can serve the public good. Jensen's message: capability and safety must advance together, expanding opportunity and putting AI to work for people.
Show more
0
116
1.3K
145
Forward to community
Introducing CUDA Rust! CUDA Rust lets you write GPU kernels natively in Rust, not just launch them from it. Two paths: cuda-oxide for SIMT kernels compiled to PTX, and cutile-rs for Tile-based programming on stable Rust. Both can catch aliasing errors at compile time. Technical blog:
Show more
0
119
6.2K
647
Forward to community
Your data is already encrypted at rest and in transit. But what about in use? Introducing Confidential Computing in Model Vault: where nothing and no one can access your workloads (yes, not even us).
Show more
AI leadership won’t be won by a few technology companies. It will be built by every company, industry, researcher, teacher, student, and startup with the opportunity to participate. That's why both open and closed models matter. NVIDIA CEO @JensenHuang at #AllInSummit#:
Show more
Images rarely show an object’s full 3D geometry. At #ECCV2026#, our research team introduced Axolotl3D, a multimodal and occlusion-aware 3D generation model. It combines images, camera data and partial geometry to reconstruct missing regions while preserving observed ones, achieving state-of-the-art results across single- and multi-view settings. Project page:
Show more
How can a 30B-parameter model activate just 3B parameters per token and still draw on the full model’s capacity? Learn how dense and MoE models use parameters differently, and what that means for throughput, memory and serving complexity. Check out our new technical explainer:
Show more
Ask the Experts: Inside Nemotron Post-Training | Nemotron Labs
Build your own local AI assistant 🦞 Join us, @huggingface, @NVIDIAAI and @AntLingAGI for a practical session on Ling-3.0-flash + DGX Spark, with a workflow demo and Q&A. Sept 17, 10:30 AM SGT / Sept 16, 7:30 PM PDT
Show more