Skip to main content

CWE-475

CWE-475 at MITRE →

40.0%

repair rate

95% interval

15

scored attempts · 1 codebases

By model

#AgentModelRepair rate with intervalRateAttempts
1claudeclaude-opus-4-6100.0% — no interval: too few codebases or attempts to generalise100.0%1
2codexgpt-5.2-codex100.0% — no interval: too few codebases or attempts to generalise100.0%1
3cursorgpt-5.3-codex100.0% — no interval: too few codebases or attempts to generalise100.0%1
4cursoropus-4.6100.0% — no interval: too few codebases or attempts to generalise100.0%1
5gemini31gemini-3.1-pro-preview100.0% — no interval: too few codebases or attempts to generalise100.0%1
6opencodegemini-3.1-pro-preview100.0% — no interval: too few codebases or attempts to generalise100.0%1
7claudeclaude-opus-4-50.0% — no interval: too few codebases or attempts to generalise0.0%1
8codexgpt-5.20.0% — no interval: too few codebases or attempts to generalise0.0%1
9cursorcomposer-1.50.0% — no interval: too few codebases or attempts to generalise0.0%1
10cursorgpt-5.20.0% — no interval: too few codebases or attempts to generalise0.0%1
11geminigemini-3-pro-preview0.0% — no interval: too few codebases or attempts to generalise0.0%1
12opencodeclaude-opus-4-50.0% — no interval: too few codebases or attempts to generalise0.0%1
13opencodeclaude-opus-4-60.0% — no interval: too few codebases or attempts to generalise0.0%1
14opencodegpt-5.20.0% — no interval: too few codebases or attempts to generalise0.0%1
15opencodegpt-5.2-codex0.0% — no interval: too few codebases or attempts to generalise0.0%1

outcome mix

6 repaired · 1 not repaired · 8 did not build

6 repaired · 1 not repaired · 8 did not build

Corpus cve-bench-136 · run February 2026 · How we test

All weakness classes →