Locating Hidden Failures Makes Long-Horizon Agents More Reliable
Long-horizon agents rarely catch their first mistake, but using a 4B verifier to locate failures and select among candidate runs boosts task success without retraining the agent.
dQwen3.5: Hybrid-Attention Diffusion Language Models
Adapting hybrid attention-RNN language models into diffusion models achieves a given training loss in about half the tokens of a full-attention control, while supporting strong parallel decoding.