Terminal-Bench-Science: Evaluating AI agents on scientific research workflows
Terminal-Bench · since 2026
Benchmark for evaluating AI agents on real scientific research workflows in terminal environments.
| Pricing | Free |
|---|---|
| Level | Advanced |
| Category | Science & Biotech |
| Best for | AI researchers and ML engineers |
Tags: benchmark, ai agents, science, evaluation, research
Visit Terminal-Bench-Science: Evaluating AI agents on scientific research workflows
Alternatives to Terminal-Bench-Science: Evaluating AI agents on scientific research workflows
- Hex — Collaborative analytics notebooks with AI (Magic)
- NotebookLM — Source-grounded research notebook with Audio Overviews
- Perplexity — AI answer engine with citations and Deep Research
- AlphaFold Server — Predict protein and biomolecule structures
- Chai Discovery — AI models for molecular structure prediction
- Cradle — AI-guided protein engineering platform