github Actively maintained

harbor-framework/terminal-bench-1

A benchmark for LLMs on complicated tasks in the terminal

1 awesome list

Quick read

Stars
2,564
Forks
567
Open issues
319
Commits
904

Activity and growth

Latest capture 2026-09-06 10:53

Stars · last 7 days
No history
Commits · last 7 days
No history
Stars since tracking
0
Stored snapshots
1

Classification

Metadata

Language
Python
License
Apache-2.0
Default branch
main
Created
2025-01-17
First commit
2025-01-17
Last pushed
2026-07-11
GitHub updated
2026-09-05
Last synced
2026-09-06 10:53
Stack scanned
2026-09-06 10:53
Archived
No

AI development signals

5 paths

Agent instructions and tool configuration found in this repository.

Growth history

Tracked growth

1 observed capture since 2026-09-06. Observed captures are shown by default.

Stars from first capture 0

Chart data

Observed captures only

Time horizon

All tracked data

Custom date range

Stars history

Observed snapshots

Commits history

Observed snapshots

Similar repositories

Nearest indexed repositories by embedding similarity.

THUDM/AgentBench

A Comprehensive Benchmark to Evaluate LLMs as Agents (ICLR'24)

3,726 stars
Python 4 awesome lists

google/BIG-bench

Beyond the Imitation Game collaborative benchmark for measuring and extrapolating the capabilities of language models

3,249 stars
Python 2 awesome lists