AI can generate a patch. That doesn’t mean it fixed the vulnerability. In our new Off-by-1 Labs research, 53.9% of analyzed LLM-generated patches failed to resolve the vulnerability, introduced a new one, or both. Explore the paper, dataset, and tooling:
顯示更多