github Actively maintained

raullenchai/Rapid-MLX

The fastest local AI engine for Apple Silicon. 4.2x faster than Ollama, 0.08s cached TTFT, 100% tool calling. 17 tool parsers, prompt cache, reasoning separation, cloud routing. Drop-in OpenAI replacement. Works with Claude Code, Cursor, Aider.

1 awesome list

Quick read

Stars
3,700
Forks
415
Open issues
55
Commits
2,359

Activity and growth

Latest capture 2026-09-08 10:55

Stars · last 7 days
No history
Commits · last 7 days
No history
Stars since tracking
0
Stored snapshots
1

Classification

Metadata

Language
Python
License
NOASSERTION
Default branch
main
Created
2026-02-25
First commit
2025-12-06
Last pushed
2026-09-08
GitHub updated
2026-09-08
Last synced
2026-09-08 10:55
Stack scanned
2026-09-08 10:55
Archived
No

Growth history

Tracked growth

1 observed capture since 2026-09-08. Observed captures are shown by default.

Stars from first capture 0

Chart data

Observed captures only

Time horizon

All tracked data

Custom date range

Stars history

Observed snapshots

Commits history

Observed snapshots

Similar repositories

Nearest indexed repositories by embedding similarity.

Blaizzy/mlx-vlm

MLX-VLM is a package for inference and fine-tuning of Vision Language Models (VLMs) on your Mac using MLX.

5,467 stars
Python 1 awesome list

OpenBMB/VoxCPM

VoxCPM2: Tokenizer-Free TTS for Multilingual Speech Generation, Creative Voice Design, and True-to-Life Cloning

36,913 stars
Python 2 awesome lists

FareedKhan-dev/kimi-k3-in-c

A 2.78-trillion-parameter Kimi K3 running inference on a single CPU in 8.24 GB of RAM. Portable C99: no BLAS, no framework, no GPU.

7,128 stars
C 1 awesome list

Scottcjn/legend-of-elya-n64

Legend of Elya — N64 game with a real 6.36M-parameter ternary transformer on the VR4300 MIPS III CPU. Zelda-style dungeon, AI NPCs, byte-level inference at 1.23 tok/s scalar / 2.19 tok/s on the RSP overlay (measured under ares, never on silicon). Built with libdragon.

128 stars
C 2 awesome lists

PromtEngineer/localGPT

Chat with your documents on your local device using GPT models. No data leaves your device and 100% private.

22,200 stars
Python 3 awesome lists

rohitg00/agentmemory

#1 Persistent memory for AI coding agents based on real-world benchmarks

28,137 stars
TypeScript 5 awesome lists