xLLM-AI/xllm
A high-performance inference engine for LLM, VLM, DiT and REC models, optimized for diverse AI accelerators. It is hosted in OpenAtom Foundation.
Fast, Flexible and Portable Structured Generation
Appears on
Quick read
Latest capture 2026-08-03 03:06
0 paths
Agent instructions and tool configuration found in this repository.
No config files detected.
6 observed captures since 2026-05-25. Observed captures are shown by default.
Stars from first capture +106
Observed captures only
All tracked data
Observed snapshots
Observed snapshots
Nearest indexed repositories by embedding similarity.
A high-performance inference engine for LLM, VLM, DiT and REC models, optimized for diverse AI accelerators. It is hosted in OpenAtom Foundation.
Universal LLM Deployment Engine with ML Compilation
A high-performance inference engine for LLM, VLM, DiT and REC models, optimized for diverse AI accelerators.
SGLang is a high-performance serving framework for large language models and multimodal models.
llama.cpp fork with additional SOTA quants and improved performance
Democratizing Reinforcement Learning for LLMs