๊ฐ€์ž… ํ›„ ์ดˆ๋Œ€ ๋งํฌ๋ฅผ ๊ณต์œ ํ•˜๋ฉด ๋™์˜์ƒ ์žฌ์ƒ ๋ฐ ์ดˆ๋Œ€ ๋ณด์ƒ์„ ๋ฐ›์„ ์ˆ˜ ์žˆ์Šต๋‹ˆ๋‹ค.

Andrew Carr ๐Ÿคธ
@andrew_n_carr
co-founder leading science @getcartwheel co-founder advisor @arcade_ai Past: Codex @OpenAI, Brain @GoogleAI, world ranked Tetris player
๊ฐ€์ž… July 2015
5.2K ํŒ”๋กœ์ž‰ ์ค‘    28.7K ํŒฌ
this figure from the deepseek v4.1 flash paper is a must study. look at the extremely sharp increase in quality after extending context length to 1M tokens. you see a similar gain in the mimo-v2.6 rl graphs as well. agents are context hungry
๋” ๋ณด๊ธฐ