Tip for your @NousResearch Hermes Agent Bots:
Create an "Orchestrator" Bot and then have it design its own team to optimally fulfill any/all user requests. Here's El Jefe (my orchestrator) completing a very solid roster including comprehensive soul.md entries for each Bot!
Qwen3.8-27B NVFP4 + MTP is live for GB10 🚀
Quantized with @NVIDIAAI ModelOpt 0.46.0rc1 using its shipped qwen3_5 recipe.
2.45× c1 speedup vs AR
84.3 tok/s at c8
262K context, 8/8 NIAH
17/17 zero-error runs
Model:
Repo:
Nous Portal + Buzz = 4 Agents on a team, powered by FREE models in the Nous Portal, debating which is using the best model 😂
If you're managing multiple agents, @NousResearch + @blocks Buzz has to be one of the best options out there right now!
If you SSH from your phone, I highly recommend trying Moshi!
This has replaced Termius on all my
mobile devices and undoubtedly increases my productivity while on the go 🤩
If you have a GB10 (DGX Spark) and a Hermes Agent, have a look at what Mike did! Pretty smart 🧠🧠🧠
"/learn the DGX Spark playbooks found here and produce a skill that demonstrates a complete mastery of GB10 operations"
@NVIDIAAI just gifted us a 75B MoE 🤩🤩🤩
nvidia/NVIDIA-Nemotron-Labs-3-Puzzle-75B-A9B-NVFP4
75.3B total / 9.3B active compressed from Nemotron-3-Super-120B using the Iterative Puzzle framework.
1M token context support!
Perfect for your single GB10 ♥️
Deepseek V4 Flash is one of the best "worker" models out there IMHO so this is very cool!
Yet ANOTHER free model available through the Nous Portal 😍😍😍
Can't believe people are still playing with toy lobsters and missing out on literal MONTHS of free inference through @NousResearch