antimatter15/alpaca.cpp
Locally run an Instruction-Tuned Chat-Style LLM
Run LLaMA (and Stanford-Alpaca) inference on Apple Silicon GPUs.
Appears on
Quick read
Latest capture 2026-08-15 03:05
0 paths
Agent instructions and tool configuration found in this repository.
No config files detected.
6 observed captures since 2026-05-27. Observed captures are shown by default.
Stars from first capture -6
Observed captures only
All tracked data
Observed snapshots
Observed snapshots
Nearest indexed repositories by embedding similarity.
Locally run an Instruction-Tuned Chat-Style LLM
Easy and Efficient Finetuning LLMs. (Supported LLama, LLama2, LLama3, Qwen, Baichuan, GLM , Falcon) 大模型高效量化训练+部署.
LLM as a Chatbot Service
llama.cpp fork with additional SOTA quants and improved performance
llama.go is like llama.cpp in pure Golang!
QLoRA: Efficient Finetuning of Quantized LLMs