๊ฐ€์ž… ํ›„ ์ดˆ๋Œ€ ๋งํฌ๋ฅผ ๊ณต์œ ํ•˜๋ฉด ๋™์˜์ƒ ์žฌ์ƒ ๋ฐ ์ดˆ๋Œ€ ๋ณด์ƒ์„ ๋ฐ›์„ ์ˆ˜ ์žˆ์Šต๋‹ˆ๋‹ค.

Medical Sphere
@MedicalSphereAI
The global community for advancing AI in healthcare Tag @AskMedSphere to test AI models on medical cases
๊ฐ€์ž… September 2025
1 ํŒ”๋กœ์ž‰ ์ค‘    2.2K ํŒฌ
We benchmarked Muse Spark 1.1 and GPT-5.6 Sol on HealthBench Professional, OpenAI's benchmark of 525 real clinician tasks ๐Ÿฅ๐Ÿฉบ Muse Spark 1.1 tops our board: better overall score than GPT-5.6 Sol, statistically on par on the length-adjusted score at a fraction of the cost ($1.25/$4.25 vs $5/$30 per M tokens in/out, ~7ร— cheaper on output).
๋” ๋ณด๊ธฐ