Terminal-Bench 2.1 Extended
AI benchmark · Agentic
- Domains
- Agentic
- Data source
- benchlm-benchmarks
- Snapshot
Full listing name: Terminal-Bench 2.1 Extended (Mercor)
About Terminal-Bench 2.1 Extended
Completing tasks in command-line environments. This table shows Mercor-run configurations for reference and is excluded from model rankings.
Where Terminal-Bench 2.1 Extended sits
Its domains, and the nearest entries sharing them. Click any node to open its page.
Select a node to trace its connections. Zoom in for more space; drag to pan.
Next: browse all benchmarks. Entry from the RL Research daily scrape of public sources, 2026-10-02 snapshot.