We're working on making the local model experience better in Hermes, what are the best local models at each weight class?
My blind guess, please correct:
8-16 GB VRAM
Gemma4 12B
24-32 GB VRAM
Qwen3.6 27B
Qwen3.6 35B
128 GB VRAM (Spark, M3 Max)
??? Can you do DSv4-Flash?