github Actively maintained

xLLM-AI/xllm

A high-performance inference engine for LLM, VLM, DiT and REC models, optimized for diverse AI accelerators. It is hosted in OpenAtom Foundation.

0 awesome lists

Quick read

Stars
1,555
Forks
290
Open issues
210
Commits
1,416

Activity and growth

Latest capture 2026-09-04 03:03

Stars · last 7 days
+15 +1.0%
Commits · last 7 days
+9 +0.6%
Stars since tracking
+85
Stored snapshots
45

Classification

Metadata

Language
C++
License
Apache-2.0
Default branch
main
Created
2025-08-12
First commit
2025-08-13
Last pushed
2026-09-04
GitHub updated
2026-09-04
Last synced
2026-09-04 03:03
Stack scanned
2026-09-04 03:03
Archived
No

Growth history

Tracked growth

45 observed captures since 2026-07-15. Observed captures are shown by default.

Stars from first capture +85

Chart data

Observed captures only

Time horizon

All tracked data

Custom date range

Stars history

Observed snapshots

Commits history

Observed snapshots

Similar repositories

Nearest indexed repositories by embedding similarity.

jd-opensource/xllm

A high-performance inference engine for LLM, VLM, DiT and REC models, optimized for diverse AI accelerators.

1,370 stars
C++ 1 awesome list

vllm-project/vllm

A high-throughput and memory-efficient inference and serving engine for LLMs

90,045 stars
Python 4 awesome lists

llm-d/llm-d

Achieve state of the art inference performance with modern accelerators on Kubernetes

3,959 stars
Shell 2 awesome lists

alibaba/rtp-llm

RTP-LLM: Alibaba's high-performance LLM inference engine for diverse applications.

1,292 stars
Cuda 1 awesome list

mlc-ai/mlc-llm

Universal LLM Deployment Engine with ML Compilation

23,020 stars
Python 2 awesome lists

sgl-project/sglang

SGLang is a high-performance serving framework for large language models and multimodal models.

31,109 stars
Python 5 awesome lists