Agent configuration · SWE-bench Verified
Atlassian Rovo Dev (2025-09-02)
The ranked object: model, scaffold, tools, permissions, budget, and environment together.
System identity
Agent systemAtlassian Rovo Dev (2025-09-02)
Base modelClaude Sonnet 4 + GPT-5
Source model ID or labelclaude-sonnet-4-20250514; gpt-5
ScaffoldAtlassian Rovo Dev
Scaffold version2025-09-02
ToolsNot reported
Comparability groupswe-bench-verified-full-500
Evaluation configuration
Source snapshotSWE-bench Verified · Verified · Full leaderboard · 500 tasks · Jul 30, 2026
Tasks500 human-validated GitHub issues
EnvironmentReproducible repository environments evaluated by SWE-bench
BudgetLimits vary by submitted configuration
Evaluation date
Data captured
Resolved76.8%
Partial scoreNot reported
Published evidence
Open sourceEvidenceListed by the benchmark publisher
SourceSWE-bench Verified
CostNot reported
Cost basisThe source does not report a comparable cost for this configuration
Comparable cost is not reported for this configuration, so it is not treated as a low-cost option.
What this result does not prove
This result applies to the exact system and source conditions shown here. It does not establish a universal best agent, production reliability, data governance, or performance on a different benchmark version.