Multi-SWE Bench
AI benchmark · Coding
- Domains
- Coding
- Data source
- benchlm-benchmarks
- Snapshot
About Multi-SWE Bench
A multi-language software-engineering benchmark that measures repository-level bug fixing and implementation across more than one programming ecosystem.
Where Multi-SWE Bench sits
Its domains, and the nearest entries sharing them. Click any node to open its page.
Select a node to trace its connections. Zoom in for more space; drag to pan.
Next: browse all benchmarks. Entry from the RL Research daily scrape of public sources, 2026-09-21 snapshot.