๊ฐ€์ž… ํ›„ ์ดˆ๋Œ€ ๋งํฌ๋ฅผ ๊ณต์œ ํ•˜๋ฉด ๋™์˜์ƒ ์žฌ์ƒ ๋ฐ ์ดˆ๋Œ€ ๋ณด์ƒ์„ ๋ฐ›์„ ์ˆ˜ ์žˆ์Šต๋‹ˆ๋‹ค.

Qwen
@Alibaba_Qwen
Open foundation models for AGI.
๊ฐ€์ž… February 2024
3 ํŒ”๋กœ์ž‰ ์ค‘    291.5K ํŒฌ
CommerceAgentBench starts with real commercial demand, and Qwen3.8-Max delivers the strongest overall performance among open-weight models. Let's test Qwen on your real-world workflows! ๐Ÿ”ฅ
Most AI benchmarks test what a model says. In commerce, the hard part was never the answer. Itโ€™s execution. Weโ€™ve open-sourced CommerceAgentBench: a benchmark for real commerce operations. Early results are humbling. The best overall completion rate is ~62%. Qwen @Alibaba_Qwen delivered the strongest overall performance across complex commercial workflows among the open-weight models evaluated. Explore the benchmark and full results โ†“
๋” ๋ณด๊ธฐ