๊ฐ€์ž… ํ›„ ์ดˆ๋Œ€ ๋งํฌ๋ฅผ ๊ณต์œ ํ•˜๋ฉด ๋™์˜์ƒ ์žฌ์ƒ ๋ฐ ์ดˆ๋Œ€ ๋ณด์ƒ์„ ๋ฐ›์„ ์ˆ˜ ์žˆ์Šต๋‹ˆ๋‹ค.

Yuhang Yao
@yuhang_yao
Senior Research Scientist @ZOOM | PhD @CarnegieMellon | Graph Agent Research
๊ฐ€์ž… December 2021
117 ํŒ”๋กœ์ž‰ ์ค‘    1K ํŒฌ
Grok 4.6 just topped the VISTA Leaderboard ๐Ÿ‘‘ ๐Ÿฅ‡ Grok 4.6 โ€” 0.552 ๐Ÿฅˆ GPT-5.6-sol โ€” 0.538 ๐Ÿฅ‰ fable-5 โ€” 0.533 VISTA tests how well coding agents turn Figma designs into functional web appsโ€”not traditional coding problems. Grok 4.6 improves significantly over 4.5โ€™s 0.517. One caveat: 4.6 ran at High effort, while 4.5 ran at Medium, so the longer runtime is expected. AI coding is moving beyond โ€œdoes it run?โ€ to โ€œdoes it look right and actually work?โ€ Models are shipping faster than ever. Looking forward to Grok 4.7โ€”and more challengers. ๐Ÿš€ #Grok46# #AICoding# #VISTA#
๋” ๋ณด๊ธฐ