๊ฐ€์ž… ํ›„ ์ดˆ๋Œ€ ๋งํฌ๋ฅผ ๊ณต์œ ํ•˜๋ฉด ๋™์˜์ƒ ์žฌ์ƒ ๋ฐ ์ดˆ๋Œ€ ๋ณด์ƒ์„ ๋ฐ›์„ ์ˆ˜ ์žˆ์Šต๋‹ˆ๋‹ค.

Mert รœnsal
@mertunsal2020
Training models and building infra @MistralAI, prev. founding engineer @browser_use (YC W25), Kimina Prover @ProjectNumina @ETH_en
๊ฐ€์ž… January 2018
1.1K ํŒ”๋กœ์ž‰ ์ค‘    3.4K ํŒฌ
Leanstral ahead of GPT-6-Astra on ArxivLean ๐Ÿš€ we used a heavily parallel multi-agent scaffold and itโ€™s with a ton of more compute and I am pretty sure Astra would do better if you scale its compute as well. but we do what we can!
๋” ๋ณด๊ธฐ
A lot of grinding from Leanstral 1.5 with a tiny bit of K3 orchestration get to beat GPT-6 Astra on ArXivLean. I think #tokens# is counted erroneously here but yes a 6B model does yap. Great internship work by Matรฉo Pirio Rossignol mentored by @roman_soletskyi!
๋” ๋ณด๊ธฐ