github Actively maintained

mostlygeek/llama-swap

Reliable model swapping for any local OpenAI/Anthropic compatible server - llama.cpp, vllm, etc

Quick read

Stars
5,247
Forks
399
Open issues
68
Commits
545

Activity and growth

Latest capture 2026-08-03 03:07

Stars · last 7 days
No history
Commits · last 7 days
No history
Stars since tracking
+988
Stored snapshots
6

Classification

Metadata

Language
Go
License
MIT
Default branch
main
Created
2024-10-04
First commit
2024-10-04
Last pushed
2026-08-03
GitHub updated
2026-08-03
Last synced
2026-08-03 03:07
Stack scanned
2026-08-03 03:07
Archived
No

AI development signals

3 paths

Agent instructions and tool configuration found in this repository.

Agent instructions

Claude Code

CodeRabbit

View paths

Growth history

Tracked growth

6 observed captures since 2026-05-25. Observed captures are shown by default.

Stars from first capture +988

Chart data

Observed captures only

Time horizon

All tracked data

Custom date range

Stars history

Observed snapshots

Commits history

Observed snapshots

Similar repositories

Nearest indexed repositories by embedding similarity.

gotzmann/llama.go

llama.go is like llama.cpp in pure Golang!

1,398 stars
Go 2 awesome lists

b4rtaz/distributed-llama

Distributed LLM inference. Connect home devices into a powerful cluster to accelerate LLM inference. More devices means faster inference.

3,019 stars
C++ 2 awesome lists

defilantech/LLMKube

Kubernetes operator for self-hosted LLM inference across a heterogeneous GPU fleet: NVIDIA CUDA, AMD Vulkan, and Apple Silicon Metal. Runtimes: llama.cpp, vLLM, TGI, mlx-server. Multi-GPU sharding, model caching, OpenAI-compatible endpoints. Apache-2.0, run across homelab and on-prem fleets, actively developed.

196 stars
Go 1 awesome list

getumbrel/llama-gpt

A self-hosted, offline, ChatGPT-like chatbot. Powered by Llama 2. 100% private, with no data leaving your device. New: Code Llama support!

10,936 stars
TypeScript 3 awesome lists