github Actively maintained

ulab-uiuc/LLMRouter

LLMRouter: An Open-Source Library for LLM Routing

1 awesome list

Quick read

Stars
2,812
Forks
302
Open issues
40
Commits
464

Activity and growth

Latest capture 2026-09-10 10:55

Stars · last 7 days
0 0.0%
Commits · last 7 days
0 0.0%
Stars since tracking
0
Stored snapshots
1

Classification

Technology stack

Metadata

Language
Python
License
MIT
Default branch
main
Created
2025-10-07
First commit
2025-10-11
Last pushed
2026-09-09
GitHub updated
2026-09-10
Last synced
2026-09-10 10:55
Stack scanned
2026-09-10 10:55
Archived
No

AI development signals

0 paths

Agent instructions and tool configuration found in this repository.

No config files detected.

Growth history

Tracked growth

1 observed capture since 2026-09-10. Observed captures are shown by default.

Stars from first capture 0

Chart data

Observed captures only

Time horizon

All tracked data

Custom date range

Stars history

Observed snapshots

Commits history

Observed snapshots

Similar repositories

Nearest indexed repositories by embedding similarity.

lm-sys/RouteLLM

A framework for serving and evaluating LLM routers - save LLM costs without compromising quality

5,457 stars
Python 1 awesome list

aurelio-labs/semantic-router

Superfast AI decision making and intelligent processing of multi-modal data.

3,872 stars
Python 1 awesome list

huggingface/chat-ui

The open source codebase powering HuggingChat

10,893 stars
TypeScript 4 awesome lists

NadirRouter/NadirClaw

Open-source LLM router & AI cost optimizer. Routes simple prompts to cheap/local models, complex ones to premium — automatically. Drop-in OpenAI-compatible proxy for Claude Code, Codex, Cursor, OpenClaw. Saves 40-70% on AI API costs. Self-hosted, no middleman.

652 stars
Python 1 awesome list

decolua/9router

Unlimited FREE AI coding. Connect Claude Code, Codex, Cursor, Cline, Copilot, Antigravity to FREE Claude/GPT/Gemini via 40+ providers. Auto-fallback, RTK -40% tokens, never hit limits.

25,511 stars
JavaScript 2 awesome lists

defilantech/LLMKube

Kubernetes operator for self-hosted LLM inference across a heterogeneous GPU fleet: NVIDIA CUDA, AMD Vulkan, and Apple Silicon Metal. Runtimes: llama.cpp, vLLM, TGI, mlx-server. Multi-GPU sharding, model caching, OpenAI-compatible endpoints. Apache-2.0, run across homelab and on-prem fleets, actively developed.

196 stars
Go 1 awesome list