๊ฐ€์ž… ํ›„ ์ดˆ๋Œ€ ๋งํฌ๋ฅผ ๊ณต์œ ํ•˜๋ฉด ๋™์˜์ƒ ์žฌ์ƒ ๋ฐ ์ดˆ๋Œ€ ๋ณด์ƒ์„ ๋ฐ›์„ ์ˆ˜ ์žˆ์Šต๋‹ˆ๋‹ค.

ModelScope
@ModelScope2022
Driving innovations with open communities. ๐Ÿ’ฌ Join our Discord:
๊ฐ€์ž… April 2024
183 ํŒ”๋กœ์ž‰ ์ค‘    16K ํŒฌ
๐Ÿš€ ZDTaichu5.0-9B is now on ModelScope! ๐Ÿค– An on-device multimodal model from TaichuAI. At 9B parameters it runs on a single GPU and brings spatial reasoning, embodied AI and agentic tool use to edge deployment. Qwen3.5-9B backbone + C-RADIOv4-H vision encoder, 128K context, any-resolution image and video input. ๐Ÿงญ Spatial reasoning: leads the compared 10B-scale open VLMs (Qwen3.5-9B, STEP3-VL-10B, gemma4-8B-E4B) and scores above Gemini 3 Pro, Grok 4 and GPT-5.2 on ViewSpatial, MMSI-Bench and MindCube-tiny ๐Ÿ› ๏ธ Agent: highest among the compared open models on TAU2-Bench, Claw-Eval and IFEval ๐Ÿ“„ First-tier results on documents, charts, OCR, visual math and video, with a ready-to-use vLLM branch and Docker image ๐Ÿง  Entropy-Gated Adaptive Recurrent Reasoning: extra latent refinement steps go only to the hard tokens
๋” ๋ณด๊ธฐ