LEADERBOARD
ANNOUNCEMENT
RUN
CONTRIBUTE
COMMUNITY
CONTRIBUTORS
TERMINAL-BENCH-
SCIENCE
0.
0
9
8
7
6
5
4
3
2
1
0
A benchmark for evaluating AI agents on research workflows across scientific domains
Run the benchmark
View the tasks
Contribute a task
Cite the benchmark