github Actively maintained

intel/neural-compressor

SOTA low-bit LLM quantization (INT8/FP8/MXFP8/INT4/MXFP4/NVFP4) & sparsity; leading model compression techniques on PyTorch, TensorFlow, and ONNX Runtime

1 awesome list

Quick read

Stars
2,707
Forks
322
Open issues
25
Commits
4,134

Activity and growth

Latest capture 2026-09-06 10:54

Stars · last 7 days
No history
Commits · last 7 days
No history
Stars since tracking
0
Stored snapshots
1

Classification

Metadata

Language
Python
License
Apache-2.0
Default branch
main
Created
2020-07-21
First commit
2020-06-16
Last pushed
2026-09-05
GitHub updated
2026-09-04
Last synced
2026-09-06 10:54
Stack scanned
2026-09-06 10:54
Archived
No

AI development signals

1 path

Agent instructions and tool configuration found in this repository.

GitHub Copilot

View paths

Growth history

Tracked growth

1 observed capture since 2026-09-06. Observed captures are shown by default.

Stars from first capture 0

Chart data

Observed captures only

Time horizon

All tracked data

Custom date range

Stars history

Observed snapshots

Commits history

Observed snapshots

Similar repositories

Nearest indexed repositories by embedding similarity.

vllm-project/llm-compressor

Transformers-compatible library for applying various compression algorithms to LLMs for optimized deployment with vLLM

3,770 stars
Python 2 awesome lists

AutoGPTQ/AutoGPTQ

An easy-to-use LLMs quantization package with user-friendly apis, based on GPTQ algorithm.

5,071 stars
Python 1 awesome list

NVIDIA/Model-Optimizer

A unified library of SOTA model optimization techniques like quantization, distillation, pruning, neural architecture search, speculative decoding, etc. It compresses deep learning models for downstream deployment frameworks like TensorRT-LLM, TensorRT, vLLM, etc. to optimize inference speed.

3,780 stars
Python 2 awesome lists

pytorch/ao

PyTorch native quantization for training and inference

2,977 stars
Python 1 awesome list

qualcomm/GenieX

Run frontier LLMs and VLMs locally on Qualcomm devices across NPU, GPU, and CPU with a few lines of code

8,354 stars
Rust 1 awesome list