github Actively maintained

pinchbench/skill

PinchBench is a benchmarking system for evaluating LLM models as OpenClaw coding agents. Made with 🦀 by the humans at https://kilo.ai

1 awesome list

Quick read

Stars
1,301
Forks
151
Open issues
23
Commits
383

Activity and growth

Latest capture 2026-08-03 03:08

Stars · last 7 days
No history
Commits · last 7 days
No history
Stars since tracking
+105
Stored snapshots
6

Classification

Technology stack

Metadata

Language
Python
License
MIT
Default branch
main
Created
2026-02-11
First commit
2026-02-11
Last pushed
2026-07-02
GitHub updated
2026-07-31
Last synced
2026-08-03 03:08
Stack scanned
2026-08-03 03:08
Archived
No

Growth history

Tracked growth

6 observed captures since 2026-05-25. Observed captures are shown by default.

Stars from first capture +105

Chart data

Observed captures only

Time horizon

All tracked data

Custom date range

Stars history

Observed snapshots

Commits history

Observed snapshots

Similar repositories

Nearest indexed repositories by embedding similarity.

InternLM/WildClawBench

An in-the-wild benchmark for AI agents in the OpenClaw Environment.

498 stars
Python 1 awesome list

Agnuxo1/benchclaw

BenchClaw — Multi-dimensional AI agent evaluation with 17-judge AI Tribunal, 10 scoring dimensions, radar charts, and deception detection. Benchmark any LLM agent.

7 stars
HTML 0 awesome lists

claw-eval/claw-eval

Claw-Eval is an evaluation harness for evaluating LLM as agents. All tasks verified by humans.

739 stars
Python 2 awesome lists

TIGER-AI-Lab/ClawBench

Open-source benchmark for browser AI agents on daily tasks.

538 stars
Python 1 awesome list