ikawrakow/ik_llama.cpp
llama.cpp fork with additional SOTA quants and improved performance
Running Llama 2 and other Open-Source LLMs on CPU Inference Locally for Document Q&A
Appears on
Quick read
Latest capture 2026-08-15 03:05
0 paths
Agent instructions and tool configuration found in this repository.
No config files detected.
6 observed captures since 2026-05-27. Observed captures are shown by default.
Stars from first capture -3
Observed captures only
All tracked data
Observed snapshots
Observed snapshots
Nearest indexed repositories by embedding similarity.
llama.cpp fork with additional SOTA quants and improved performance
Inference code for Llama models
A school for camelids
LangChain & Prompt Engineering tutorials on Large Language Models (LLMs) such as ChatGPT with custom data. Jupyter notebooks on loading and indexing data, creating prompt templates, CSV agents, and using retrieval QA chains to query the custom data. Projects for using a private LLM (Llama 2) for chat with PDF files, tweets sentiment analysis.
A fast inference library for running LLMs locally on modern consumer-class GPUs
A self-hosted, offline, ChatGPT-like chatbot. Powered by Llama 2. 100% private, with no data leaving your device. New: Code Llama support!