github Actively maintained

ARBML/Taqyim

Python intefrace for evaluation on chatgpt models

1 awesome list

Quick read

Stars
19
Forks
4
Open issues
1
Commits
46

Activity and growth

Latest capture 2026-08-12 03:03

Stars · last 7 days
No history
Commits · last 7 days
No history
Stars since tracking
0
Stored snapshots
6

Metadata

Language
Jupyter Notebook
License
MIT
Default branch
main
Created
2023-05-27
First commit
2023-05-27
Last pushed
2024-02-13
GitHub updated
2024-05-17
Last synced
2026-08-12 03:03
Stack scanned
2026-08-12 03:03
Archived
No

AI development signals

0 paths

Agent instructions and tool configuration found in this repository.

No config files detected.

Growth history

Tracked growth

6 observed captures since 2026-05-27. Observed captures are shown by default.

Stars from first capture 0

Chart data

Observed captures only

Time horizon

All tracked data

Custom date range

Stars history

Observed snapshots

Commits history

Observed snapshots

Similar repositories

Nearest indexed repositories by embedding similarity.

EvolvingLMMs-Lab/lmms-eval

One-for-All Multimodal Evaluation Toolkit Across Text, Image, Video, and Audio Tasks

4,343 stars
Python 1 awesome list

claw-eval/claw-eval

Claw-Eval is an evaluation harness for evaluating LLM as agents. All tasks verified by humans.

739 stars
Python 2 awesome lists

sierra-research/tau2-bench

τ-Bench: A Benchmark for Tool-Agent-User Interaction in Real-World Domains

1,712 stars
Python 1 awesome list

google/BIG-bench

Beyond the Imitation Game collaborative benchmark for measuring and extrapolating the capabilities of language models

3,249 stars
Python 2 awesome lists