With just one prompt, we taught an LLM to see - then looked at its representations to debug where it fell short.
Qwen 3 8B can read text, but has no way to see or understand images. Silico trained a vision adapter that matches the official Qwen 3 VL 8B in multiple benchmarks. 🧵
显示更多