Agents trained on population-scale data can do a lot of things.
But ask a professional to stake their reputation on an AI-generated artifact? “Pretty good” isn’t good enough.
We introduce TAHI: a Test-time Adaptive agent framework through Human-agent Interaction. Featuring:
📈 Test-time adaptation via context (memory, skills) and weight training
⚡ Efficient adaptation to individual expertise within tens of tasks
✔️ Creating comprehensive rubrics for “non-verifiable” tasks
🔍 Analysis of shared community guidelines vs. personalized tacit expertise