One year in, the
@DARPA AIQ program is moving beyond better benchmarks toward a science of AI capability: measuring what models can do on specific questions, predicting performance across classes of problems, and understanding how capabilities arise from model architectures.