github Actively maintained

Michael-A-Kuykendall/shimmy

⚡ Pure-Rust WebGPU inference engine — OpenAI-API compatible, GGUF native, runs on any GPU. No Python. No llama.cpp. Single binary.

Quick read

Stars
5,727
Forks
550
Open issues
10
Commits
135

Activity and growth

Latest capture 2026-08-03 03:09

Stars · last 7 days
No history
Commits · last 7 days
No history
Stars since tracking
+461
Stored snapshots
6

Classification

Metadata

Language
Rust
License
Apache-2.0
Default branch
main
Created
2025-08-28
First commit
2026-01-10
Last pushed
2026-08-02
GitHub updated
2026-08-03
Last synced
2026-08-03 03:09
Stack scanned
2026-08-03 03:09
Archived
No

Growth history

Tracked growth

6 observed captures since 2026-05-25. Observed captures are shown by default.

Stars from first capture +461

Chart data

Observed captures only

Time horizon

All tracked data

Custom date range

Stars history

Observed snapshots

Commits history

Observed snapshots

Similar repositories

Nearest indexed repositories by embedding similarity.

Farama-Foundation/Shimmy

PettingZoo and Gymnasium bindings for popular reinforcement learning environments outside of Farama

222 stars
Python 1 awesome list

EricLBuehler/candle-vllm

Efficent platform for inference and serving local LLMs including an OpenAI compatible API server.

712 stars
Rust 1 awesome list

kaito-project/aikit

🏗️ Fine-tune, build, and deploy open-source LLMs easily!

535 stars
Go 2 awesome lists

lyogavin/airllm

AirLLM 70B inference with single 4GB GPU

28,738 stars
Jupyter Notebook 3 awesome lists