jd-opensource/xllm
A high-performance inference engine for LLM, VLM, DiT and REC models, optimized for diverse AI accelerators.
A high-performance inference engine for LLM, VLM, DiT and REC models, optimized for diverse AI accelerators. It is hosted in OpenAtom Foundation.
Quick read
Latest capture 2026-09-04 03:03
37 paths
Agent instructions and tool configuration found in this repository.
Agent instructions
Agent workspace 27
Claude Code 3
Cursor 3
Gemini CLI 3
13 more paths detected.
45 observed captures since 2026-07-15. Observed captures are shown by default.
Stars from first capture +85
Observed captures only
All tracked data
Observed snapshots
Observed snapshots
Nearest indexed repositories by embedding similarity.
A high-performance inference engine for LLM, VLM, DiT and REC models, optimized for diverse AI accelerators.
A high-throughput and memory-efficient inference and serving engine for LLMs
Achieve state of the art inference performance with modern accelerators on Kubernetes
RTP-LLM: Alibaba's high-performance LLM inference engine for diverse applications.
Universal LLM Deployment Engine with ML Compilation
SGLang is a high-performance serving framework for large language models and multimodal models.