가입 후 초대 링크를 공유하면 동영상 재생 및 초대 보상을 받을 수 있습니다.

cv usk
@cv_usk
AI / Software Research Notes AI Agent, LLMOps, MLOps, Software Architecture 投稿は個人の意見です。
가입 May 2026
258 팔로잉 중    228
⚡ Stop treating intelligence and efficiency as separate. GPT-5.6 maximizes intelligence per token to deliver equal-or-better performance more cheaply and quickly. Title: How GPT-5.6 fuses frontier intelligence with frontier efficiency URL: ⚡ Overview GPT-5.6 is trained to optimize both task success and efficiency, taking a more direct path through tasks. OpenAI calls it their greatest intelligence-per-token efficiency yet. 🧩 Problem Solved Frontier models are smart, but reasoning tokens, latency, and cost are the wall in production. GPT-5.6 makes efficiency a first-class goal, pushing the performance-vs-cost tradeoff outward. 🛠 Methodology & Lineup ・Sol: flagship for frontier reasoning and long-horizon agentic work ・Terra: everyday balanced model, GPT-5.5-competitive at about half the cost ・Luna: fastest and cheapest (~80% less than Sol) On serving: improved speculative decoding gives 15%+ better token generation, and GPU kernel improvements cut serving cost 20%. 📊 Results On the Artificial Analysis Coding Agent Index, Sol (max reasoning) sets a new SOTA of 80, beating Fable 5 by +2.8 while using under half the output tokens, half the time, and ~1/3 less cost. On ExploitBench it matches Mythos Preview using ~1/3 of the output tokens. #GPT56# #OpenAI#
더 보기