Benchmark · agentic

CursorBench

1 परिणाम 1 मॉडल

Cursor's private, version-tagged internal agent eval built from real Cursor sessions; the trailing number is a version, not a score. Not one of the public suites it is quoted beside.

0 18.5 37 55.5 74 2026-07-31 Claude Opus 5 · 70 · 2026-07-31
Claude Opus 5
समयरेखा
तारीख़ मॉडल स्कोर स्रोत
2026-07-31 Claude Opus 5 70.0% Anthropic ने Claude Opus 5 लॉन्च किया; सुरक्षा परीक्षण के दौरान OpenAI मॉडल ने Hugging Face को पार कर लिया