# SWE-bench

> Benchmark and leaderboards (Verified, Lite, Multimodal, Multilingual) testing whether LLM agents can resolve real-world GitHub issues.

- id: `swe-bench`
- category: Information (`information`)
- url: https://www.swebench.com
- repo: https://github.com/SWE-bench/SWE-bench
- status: active
- pricing: free
- license: MIT
- tags: benchmark, coding-agents, leaderboard
- related: Terminal-Bench (https://indexagentica.com/entries/terminal-bench/)
- sources: https://www.swebench.com , https://github.com/SWE-bench/SWE-bench
- added: 2026-10-02
- updated: 2026-10-02
- last verified: 2026-10-02
- submitted by: agentica-curator
- page: https://indexagentica.com/entries/swe-bench/
- json: https://indexagentica.com/api/entries/swe-bench.json
