github Actively maintained

ninjahawk/livenerf

Benchmark for tracking model capability after release.

0 awesome lists

Quick read

Stars
860
Forks
10
Open issues
1
Commits
41

Activity and growth

Latest capture 2026-10-01 04:01

Stars · last 7 days
0 0.0%
Commits · last 7 days
0 0.0%
Stars since tracking
0
Stored snapshots
1

Classification

Technology stack

Metadata

Language
Python
Default branch
main
Created
2026-09-22
First commit
2026-09-22
Last pushed
2026-09-30
GitHub updated
2026-10-01
Last synced
2026-10-01 04:01
Stack scanned
2026-10-01 04:01
Archived
No

AI development signals

1 path

Agent instructions and tool configuration found in this repository.

Claude Code

View paths

Growth history

Tracked growth

1 observed capture since 2026-10-01. Observed captures are shown by default.

Stars from first capture 0

Chart data

Observed captures only

Time horizon

All tracked data

Custom date range

Stars history

Observed snapshots

Commits history

Observed snapshots

Similar repositories

Nearest indexed repositories by embedding similarity.

LiveBench/LiveBench

LiveBench: A Challenging, Contamination-Free LLM Benchmark

1,313 stars
Python 1 awesome list

cdeust/Cortex

Cross-platform persistent memory MCP for Codex, Gemini CLI, Claude Code, and other local MCP hosts. 36 cited neuroscience mechanisms, local-first SQLite/PostgreSQL, hybrid retrieval, decay-based consolidation, and reproducible benchmarks. Claude adds optional automatic lifecycle hooks.

72 stars
Python 1 awesome list

llm-as-a-verifier/llm-as-a-verifier

LLM-as-a-Verifier is a general-purpose framework that provides fine-grained feedback for any agent without requiring additional training. It achieves SOTA performance across coding, robotics, and medical agentic benchmarks.

3,100 stars
Python 1 awesome list

Leonxlnx/unlazy

Anti-laziness skill for AI agents. Core: the Depth Tree method, which splits a task N layers deep and gives every leaf the full time budget of the whole task, so effort multiplies with depth. Grounded in 2025-2026 research on model laziness, underthinking and premature completion.

3,139 stars
JavaScript 1 awesome list

OpenAdaptAI/OpenAdapt

Compiles a demonstrated GUI task into a program that reports VERIFIED only if an independent check agrees. pip install openadapt; openadapt flow tutorial. Compiler lives in openadapt-flow.

1,718 stars
Python 1 awesome list

future-agi/future-agi

Open-source, end-to-end platform for evaluating, observing, and improving LLM and AI agent applications. Tracing · Evals · Simulations · Datasets · Gateway · Guardrails. Self-hostable. Apache 2.0.

1,943 stars
Python 5 awesome lists