Multi-SWE Bench

AI benchmark · Coding

Domains
Coding
Data source
benchlm-benchmarks
Snapshot

Open benchmark ↗benchlm.ai/benchmarks/multiSweBench

About Multi-SWE Bench

A multi-language software-engineering benchmark that measures repository-level bug fixing and implementation across more than one programming ecosystem.

Where Multi-SWE Bench sits

Its domains, and the nearest entries sharing them. Click any node to open its page.

Select a node to trace its connections. Zoom in for more space; drag to pan.

Next: browse all benchmarks. Entry from the RL Research daily scrape of public sources, 2026-09-21 snapshot.