ikawrakow/ik_llama.cpp
llama.cpp fork with additional SOTA quants and improved performance
Run GGUF models easily with a KoboldAI UI. One File. Zero Install.
Appears on
Quick read
Latest capture 2026-08-03 03:05
0 paths
Agent instructions and tool configuration found in this repository.
No config files detected.
6 observed captures since 2026-05-25. Observed captures are shown by default.
Stars from first capture +699
Observed captures only
All tracked data
Observed snapshots
Observed snapshots
Nearest indexed repositories by embedding similarity.
llama.cpp fork with additional SOTA quants and improved performance
🏗️ Fine-tune, build, and deploy open-source LLMs easily!
Local UI to run and train LLMs and diffusion models. Supports GGUF, MLX, Qwen3.8, DeepSeek-V4, MiniMax-H3, Gemma 4, FLUX and more.
Unified Efficient Fine-Tuning of 100+ LLMs & VLMs (ACL 2024)
A list of free LLM inference resources accessible via API.
LLM inference in C/C++