ollama/ollama
Get up and running with Kimi-K2.6, GLM-5.2, MiniMax, DeepSeek, gpt-oss, Qwen, Gemma and other models.
Small self-contained pure-Go web server with Lua, Teal, Markdown, HTTP/2, QUIC, Redis, TypeScript, React 19, SQLite, and PostgreSQL support ++
Appears on
Quick read
Latest capture 2026-08-01 03:07
1 path
Agent instructions and tool configuration found in this repository.
Agent instructions
6 observed captures since 2026-05-25. Observed captures are shown by default.
Stars from first capture +12
Observed captures only
All tracked data
Observed snapshots
Observed snapshots
Nearest indexed repositories by embedding similarity.
Get up and running with Kimi-K2.6, GLM-5.2, MiniMax, DeepSeek, gpt-oss, Qwen, Gemma and other models.
llama.cpp fork with additional SOTA quants and improved performance
Putting a brain behind `cat`🐈⬛ Integrating language models in the Unix commands ecosystem through text streams.
LLM inference in C/C++
RamaLama is an open-source developer tool that simplifies the local serving of AI models from any source and facilitates their use for inference in production, all through the familiar language of containers.
Running Llama 2 and other Open-Source LLMs on CPU Inference Locally for Document Q&A