๊ฐ€์ž… ํ›„ ์ดˆ๋Œ€ ๋งํฌ๋ฅผ ๊ณต์œ ํ•˜๋ฉด ๋™์˜์ƒ ์žฌ์ƒ ๋ฐ ์ดˆ๋Œ€ ๋ณด์ƒ์„ ๋ฐ›์„ ์ˆ˜ ์žˆ์Šต๋‹ˆ๋‹ค.

KD
@Reveur_7
PhD @Berkeley_ai | CEO @ Embodied Science Alum @CarnegieMellon | Ex. Principal SWE #Mamba4Life#
๊ฐ€์ž… December 2018
297 ํŒ”๋กœ์ž‰ ์ค‘    198 ํŒฌ
What if a robot policy weren't a neural net or a test-time chat loop, but a multi-file code repo selected from a Pareto frontier of genetically evolved candidates? RHO moves all its LLM exploration to training time, then runs that repo on scenes it was never trained on. ๐Ÿงต๐Ÿ‘‡๐Ÿฝ
๋” ๋ณด๊ธฐ