๊ฐ€์ž… ํ›„ ์ดˆ๋Œ€ ๋งํฌ๋ฅผ ๊ณต์œ ํ•˜๋ฉด ๋™์˜์ƒ ์žฌ์ƒ ๋ฐ ์ดˆ๋Œ€ ๋ณด์ƒ์„ ๋ฐ›์„ ์ˆ˜ ์žˆ์Šต๋‹ˆ๋‹ค.

Alex Ker ๐Ÿ”ญ
@thealexker
code+words @baseten | investing in frontiers & sharing my curiosities | prev @bloombergbeta @stanfordhai @neurable.
๊ฐ€์ž… May 2018
1.3K ํŒ”๋กœ์ž‰ ์ค‘    13.3K ํŒฌ
post-training is the next frontier for scaling laws if you want to understand the efficient post-training mechanics behind open models today, I summarized all the innovations for glm-5.3 in plain english, covering environment design, architecture, as well as the rl algorithm and infra behind it:
๋” ๋ณด๊ธฐ