github Actively maintained

intentee/paddler

Open-source LLM/VLM load balancer and serving platform for self-hosting LLMs (and VLMs) at scale 🏓🦙 Alternative to projects like llm-d, Docker Model Runner, etc but with less moving parts and simple deployments built around ggml ecosystem. Runs on CPU and GPU.

0 awesome lists

Quick read

Stars
1,651
Forks
95
Open issues
27
Commits
1,193

Activity and growth

Latest capture 2026-08-11 04:25

Stars · last 7 days
No history
Commits · last 7 days
No history
Stars since tracking
+67
Stored snapshots
64

Classification

Metadata

Language
Rust
License
Apache-2.0
Default branch
main
Created
2024-04-27
First commit
2024-04-27
Last pushed
2026-07-19
GitHub updated
2026-08-11
Last synced
2026-08-11 04:25
Stack scanned
2026-08-11 04:25
Archived
No

Growth history

Tracked growth

64 observed captures since 2026-06-02. Observed captures are shown by default.

Stars from first capture +67

Chart data

Observed captures only

Time horizon

All tracked data

Custom date range

Stars history

Observed snapshots

Commits history

Observed snapshots

Similar repositories

Nearest indexed repositories by embedding similarity.

getumbrel/llama-gpt

A self-hosted, offline, ChatGPT-like chatbot. Powered by Llama 2. 100% private, with no data leaving your device. New: Code Llama support!

10,936 stars
TypeScript 3 awesome lists

mudler/LocalAI

LocalAI is the open-source AI engine. Run any model - LLMs, vision, voice, image, video - on any hardware. No GPU required.

48,515 stars
Go 6 awesome lists

llm-d/llm-d

Achieve state of the art inference performance with modern accelerators on Kubernetes

3,959 stars
Shell 2 awesome lists

ludwig-ai/ludwig

Low-code framework for building custom LLMs, neural networks, and other AI models

11,746 stars
Python 3 awesome lists

lyogavin/airllm

AirLLM 70B inference with single 4GB GPU

28,738 stars
Jupyter Notebook 3 awesome lists