註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

cv usk
@cv_usk
AI / Software Research Notes AI Agent, LLMOps, MLOps, Software Architecture 投稿は個人の意見です。
加入 May 2026
280 正在關注    422 粉絲
TL;DR Self-evolving agents that write their own questions and answer them can fall into "co-cheating," where the proposer and solver quietly agree on the same mistakes. Splitting source documents to evaluate across folds fixes this and lifts performance by over 8 points. Title: False Frontiers: Diagnosing and Mitigating Co-Cheating in Self-Evolving Search Agents URL: Points 🔁 Proposer and solver share source-derived errors, letting false agreement cycle back as reward — the paper calls this "co-cheating" 📉 Standard Dr. Zero systems show 6.1% and 8.8% false-agreement mass ✂️ CrossFit splits source documents into two folds, scoring each proposer's questions with a solver trained only on the other fold 📊 CrossFit alone cuts false agreement to 3.0%/3.7%; combined with MSV it drops to 2.0%/1.7% 🚀 Average downstream Cover-EM improves by 8.8 and 8.4 points over Dr. Zero 🧩 Multi-hop tasks see the biggest gains, averaging over 10 points 💰 Compute cost rises 1.72-2.7x over baseline, though a half-budget variant still works What stands out: without auditing the evaluator's own training history, apparent progress can be an illusion. #SelfEvolvingAgents# #ReinforcementLearning#
顯示更多