๊ฐ€์ž… ํ›„ ์ดˆ๋Œ€ ๋งํฌ๋ฅผ ๊ณต์œ ํ•˜๋ฉด ๋™์˜์ƒ ์žฌ์ƒ ๋ฐ ์ดˆ๋Œ€ ๋ณด์ƒ์„ ๋ฐ›์„ ์ˆ˜ ์žˆ์Šต๋‹ˆ๋‹ค.

Shashwat Goel
@ShashwatGoel7
Training AI for Decision Making Past work: Training AI Co-scientists, ฮ”Belief-RL, Measuring Long Horizon Execution
๊ฐ€์ž… June 2020
2.3K ํŒ”๋กœ์ž‰ ์ค‘    4.1K ํŒฌ
I've released 4 quite distinct evaluations in my PhD now: code falsification, long-horizon execution, research plan generation, and now forecasting agents. GPT models have been far ahead everytime. The wide range of tasks OpenAI post-training generalizes to is just ๐ŸคŒ Never switched from codex :)
๋” ๋ณด๊ธฐ