๊ฐ€์ž… ํ›„ ์ดˆ๋Œ€ ๋งํฌ๋ฅผ ๊ณต์œ ํ•˜๋ฉด ๋™์˜์ƒ ์žฌ์ƒ ๋ฐ ์ดˆ๋Œ€ ๋ณด์ƒ์„ ๋ฐ›์„ ์ˆ˜ ์žˆ์Šต๋‹ˆ๋‹ค.

David Hendrickson
@TeksEdge
CEO & Founder | PhD | Startup Advisor | @Columbia | Author Generative Software Engineering | ๐Ÿ”” Follow for AI & Vibe Coding Tips ๐Ÿ‘‡
๊ฐ€์ž… July 2023
549 ํŒ”๋กœ์ž‰ ์ค‘    11.2K ํŒฌ
Just sitting and reflecting on this GPT-6 Astra graph. GPT-6 Astra set a new SOTA on ARC-AGI. ๐Ÿงฉ ARC-AGI-1 โ†’ 98.5% ๐Ÿง  ARC-AGI-2 โ†’ 95.0% ๐Ÿค– ARC-AGI-3 โ†’ 99.95% But the ARC-AGI-3 result has a weird twist. ARC-AGI-3 drops Astra into a completely NEW interactive environment with no instructions. It has to ... ๐Ÿ‘€ explore ๐Ÿง  figure out the rules ๐Ÿ—บ๏ธ build a world model ๐Ÿ“Œ remember what it learned ๐ŸŽฏ develop a strategy ๐Ÿ”„ adapt when it's wrong ARC tested ~500 humans to establish how efficiently people solve these environments. GPT-6 Astra with ARC's standard agent harness: 62.7% Astra using the new Provider Adapter that preserves reasoning state between requests + compacts it: 99.95% ๐Ÿคฏ Give a powerful model a persistent working state + good compaction, and suddenly the same intelligence can behave VERY differently over a long task. That's going to matter for Local AI too.
๋” ๋ณด๊ธฐ