Terminal-Bench 3.0

AI benchmark · Agentic

Domains
Agentic
Data source
benchlm-benchmarks
Snapshot

Open benchmark ↗benchlm.ai/benchmarks/frontierBench

About Terminal-Bench 3.0

A continuously maintained benchmark for difficult computer work, including coding, deep learning, finance, engineering, math, and science tasks.

Where Terminal-Bench 3.0 sits

Its domains, and the nearest entries sharing them. Click any node to open its page.

Select a node to trace its connections. Zoom in for more space; drag to pan.

Next: browse all benchmarks. Entry from the RL Research daily scrape of public sources, 2026-09-21 snapshot.