We post-trained
@NVIDIAAI Nemotron 3.5 Lightning on Legal Agent Bench with
@trajectorylabs.
Here's what we found:
1) Post-training improved agent performance from 0% to 8.3% on held-out LAB tasks, beating both Opus 4.6 and the much larger post-trained Nemotron 3 Ultra.
2) Performance improved across nine practice areas with no regressions.
3) Post-training reduced average model output from 90k to 37k tokens, increasing the model's reward-per-token by 2.4x.
Through our collaboration with NVIDIA and Trajectory we’re committed to pushing the frontier of legal intelligence and cost efficiency with open weight models.
Deep dive: