kaito-project/aikit
🏗️ Fine-tune, build, and deploy open-source LLMs easily!
Efficient Triton Kernels for LLM Training
Appears on
Quick read
Latest capture 2026-08-03 03:05
45 paths
Agent instructions and tool configuration found in this repository.
Agent instructions
Agent workspace 42
Claude Code 2
21 more paths detected.
6 observed captures since 2026-05-25. Observed captures are shown by default.
Stars from first capture +160
Observed captures only
All tracked data
Observed snapshots
Observed snapshots
Nearest indexed repositories by embedding similarity.
🏗️ Fine-tune, build, and deploy open-source LLMs easily!
A Flexible Framework for Experiencing Heterogeneous LLM Inference/Fine-tune Optimizations
Easy and Efficient Finetuning LLMs. (Supported LLama, LLama2, LLama3, Qwen, Baichuan, GLM , Falcon) 大模型高效量化训练+部署.
A unified library of SOTA model optimization techniques like quantization, distillation, pruning, neural architecture search, speculative decoding, etc. It compresses deep learning models for downstream deployment frameworks like TensorRT-LLM, TensorRT, vLLM, etc. to optimize inference speed.
Go ahead and axolotl questions
llama.cpp fork with additional SOTA quants and improved performance