github Actively maintained

jd-opensource/xllm

A high-performance inference engine for LLM, VLM, DiT and REC models, optimized for diverse AI accelerators.

1 awesome list

Quick read

Stars
1,370
Forks
240
Open issues
162
Commits
1,115

Activity and growth

Latest capture 2026-06-28 03:08

Stars · last 7 days
No history
Commits · last 7 days
No history
Stars since tracking
+70
Stored snapshots
4

Classification

Metadata

Language
C++
License
Apache-2.0
Default branch
main
Created
2025-08-12
First commit
2025-08-13
Last pushed
2026-06-27
GitHub updated
2026-06-27
Last synced
2026-06-28 03:08
Stack scanned
2026-06-28 03:08
Archived
No

Growth history

Tracked growth

4 observed captures since 2026-05-25. Observed captures are shown by default.

Stars from first capture +70

Chart data

Observed captures only

Time horizon

All tracked data

Custom date range

Stars history

Observed snapshots

Commits history

Observed snapshots

Similar repositories

Nearest indexed repositories by embedding similarity.

xLLM-AI/xllm

A high-performance inference engine for LLM, VLM, DiT and REC models, optimized for diverse AI accelerators. It is hosted in OpenAtom Foundation.

1,555 stars
C++ 0 awesome lists

vllm-project/vllm

A high-throughput and memory-efficient inference and serving engine for LLMs

90,045 stars
Python 4 awesome lists

alibaba/rtp-llm

RTP-LLM: Alibaba's high-performance LLM inference engine for diverse applications.

1,292 stars
Cuda 1 awesome list

llm-d/llm-d

Achieve state of the art inference performance with modern accelerators on Kubernetes

3,959 stars
Shell 2 awesome lists

InternLM/lmdeploy

LMDeploy is a toolkit for compressing, deploying, and serving LLMs.

7,985 stars
Python 3 awesome lists

mlc-ai/mlc-llm

Universal LLM Deployment Engine with ML Compilation

23,020 stars
Python 2 awesome lists