EricLBuehler/mistral.rs
Fast, flexible LLM inference
Medusa: Simple Framework for Accelerating LLM Generation with Multiple Decoding Heads
Appears on
Quick read
Latest capture 2026-09-05 10:53
0 paths
Agent instructions and tool configuration found in this repository.
No config files detected.
1 observed capture since 2026-09-05. Observed captures are shown by default.
Stars from first capture 0
Observed captures only
All tracked data
Observed snapshots
Observed snapshots
Nearest indexed repositories by embedding similarity.
Fast, flexible LLM inference
An open platform for training, serving, and evaluating large language models. Release repo for Vicuna and Chatbot Arena.
Multi-LoRA inference server that scales to 1000s of fine-tuned LLMs
Go ahead and axolotl questions
Official inference library for Mistral models
Large-scale LLM inference engine