github Actively maintained

InternLM/lmdeploy

LMDeploy is a toolkit for compressing, deploying, and serving LLMs.

Quick read

Stars
7,985
Forks
723
Open issues
612
Commits
2,014

Activity and growth

Latest capture 2026-08-03 03:04

Stars · last 7 days
No history
Commits · last 7 days
No history
Stars since tracking
+112
Stored snapshots
6

Classification

Metadata

Language
Python
License
Apache-2.0
Default branch
main
Created
2023-06-15
First commit
2023-06-18
Last pushed
2026-08-01
GitHub updated
2026-08-03
Last synced
2026-08-03 03:04
Stack scanned
2026-08-03 03:04
Archived
No

AI development signals

7 paths

Agent instructions and tool configuration found in this repository.

Growth history

Tracked growth

6 observed captures since 2026-05-25. Observed captures are shown by default.

Stars from first capture +112

Chart data

Observed captures only

Time horizon

All tracked data

Custom date range

Stars history

Observed snapshots

Commits history

Observed snapshots

Similar repositories

Nearest indexed repositories by embedding similarity.

vllm-project/vllm

A high-throughput and memory-efficient inference and serving engine for LLMs

90,045 stars
Python 4 awesome lists

xLLM-AI/xllm

A high-performance inference engine for LLM, VLM, DiT and REC models, optimized for diverse AI accelerators. It is hosted in OpenAtom Foundation.

1,555 stars
C++ 0 awesome lists

llm-d/llm-d

Achieve state of the art inference performance with modern accelerators on Kubernetes

3,959 stars
Shell 2 awesome lists

mlc-ai/mlc-llm

Universal LLM Deployment Engine with ML Compilation

23,020 stars
Python 2 awesome lists

jd-opensource/xllm

A high-performance inference engine for LLM, VLM, DiT and REC models, optimized for diverse AI accelerators.

1,370 stars
C++ 1 awesome list

OptimalScale/LMFlow

An Extensible Toolkit for Finetuning and Inference of Large Foundation Models. Large Models for All.

8,486 stars
Python 2 awesome lists