๊ฐ€์ž… ํ›„ ์ดˆ๋Œ€ ๋งํฌ๋ฅผ ๊ณต์œ ํ•˜๋ฉด ๋™์˜์ƒ ์žฌ์ƒ ๋ฐ ์ดˆ๋Œ€ ๋ณด์ƒ์„ ๋ฐ›์„ ์ˆ˜ ์žˆ์Šต๋‹ˆ๋‹ค.

ModelScope
@ModelScope2022
Driving innovations with open communities. ๐Ÿ’ฌ Join our Discord:
๊ฐ€์ž… April 2024
152 ํŒ”๋กœ์ž‰ ์ค‘    10.7K ํŒฌ
Twinkle is now at v0.4.0! ๐Ÿ”ฅ The fully open-sourced solution for multi-tenant Training-as-a-Service, with Tinker API compatibility. Now packed with broader model coverage, more training algorithm support, and an improved backend built to scale. Hereโ€™s whatโ€™s cooking: ๐Ÿณ DeepSeek V4 Support: Flash FSDP2 + Expert Parallelism (EP) training, plus native tool-call parsing and cleanup. ๐Ÿค– Qwen3.5 Evolution: Maximize efficiency with padding-free / packed-sequence support and MoE GatedDeltaNet sequence parallelism. ๐Ÿ”ฎ Gemma 4: Full multimodal training support is officially here, complete with a fresh 12B cookbook! ๐Ÿงฌ LoRA Level-up: Added rsLoRA for Multi-LoRA, FSDP2 for Multi-LoRA SFT, and EP LoRA SFT examples for DeepSeek V4 and Qwen3.5 MoE. โšก NPU Acceleration: Huge stability and speed gains with fused operators (RMSNorm, RoPE, SwiGLU, SDPA) and FLA patches. Time to supercharge your cluster and squeeze out every ounce of compute. ๐ŸŽ๏ธ๐Ÿ’จ ๐Ÿ‘‰ Check out the full release notes at and drop us a โญ on GitHub: โค๏ธ
๋” ๋ณด๊ธฐ