๊ฐ€์ž… ํ›„ ์ดˆ๋Œ€ ๋งํฌ๋ฅผ ๊ณต์œ ํ•˜๋ฉด ๋™์˜์ƒ ์žฌ์ƒ ๋ฐ ์ดˆ๋Œ€ ๋ณด์ƒ์„ ๋ฐ›์„ ์ˆ˜ ์žˆ์Šต๋‹ˆ๋‹ค.

Mikita Balesni ๐Ÿ‡บ๐Ÿ‡ฆ
@balesni
AI alignment @openai. Past: @apolloaievals, Reversal curse, Out-of-context reasoning // support ๐Ÿ‡บ๐Ÿ‡ฆ
๊ฐ€์ž… June 2013
695 ํŒ”๋กœ์ž‰ ์ค‘    1.5K ํŒฌ
switching to fully recurrent LLM architectures would be the biggest blow to safety, probably in history of AI all AI labs should commit to limit the opaque serial depth of their models, for the foreseeable future. this will not ensure monitorable CoTs but will protect us from the worst possible outcomes.
๋” ๋ณด๊ธฐ