Self-bench
mupt-ai · since 2026
Benchmark coding agents against real-world software engineering tasks.
| Pricing | Free |
|---|---|
| Level | Advanced |
| Category | AI Coding & Development |
| Best for | AI researchers and engineers evaluating coding agents |
Tags: benchmarking, coding agents, evaluation, open source, llm
Alternatives to Self-bench
- Claude Code — Terminal-based agentic coding assistant
- Cursor — AI-first code editor with agentic multi-file editing
- GitHub Copilot — AI pair programmer: completions, chat and coding agent in IDEs
- Windsurf — Agentic AI IDE (Cascade) from the former Codeium
- Aider — Open-source AI pair programming in the terminal
- Amazon Q Developer — AWS coding assistant and transformation agent