github Actively maintained

Luce-Org/lucebox

LLM speculative inference server for heterogeneous hardware & consumer GPUs

1 awesome list

Quick read

Stars
2,838
Forks
271
Open issues
115
Commits
1,536

Activity and growth

Latest capture 2026-09-07 10:54

Stars · last 7 days
No history
Commits · last 7 days
No history
Stars since tracking
0
Stored snapshots
1

Classification

Metadata

Language
C++
License
Apache-2.0
Default branch
main
Created
2026-04-03
First commit
2026-04-03
Last pushed
2026-09-07
GitHub updated
2026-09-07
Last synced
2026-09-07 10:54
Stack scanned
2026-09-07 10:54
Archived
No

AI development signals

1 path

Agent instructions and tool configuration found in this repository.

Codex

View paths

Growth history

Tracked growth

1 observed capture since 2026-09-07. Observed captures are shown by default.

Stars from first capture 0

Chart data

Observed captures only

Time horizon

All tracked data

Custom date range

Stars history

Observed snapshots

Commits history

Observed snapshots

Similar repositories

Nearest indexed repositories by embedding similarity.

EricLBuehler/candle-vllm

Efficent platform for inference and serving local LLMs including an OpenAI compatible API server.

712 stars
Rust 1 awesome list

ludwig-ai/ludwig

Low-code framework for building custom LLMs, neural networks, and other AI models

11,753 stars
Python 3 awesome lists

qualcomm/GenieX

Run frontier LLMs and VLMs locally on Qualcomm devices across NPU, GPU, and CPU with a few lines of code

8,354 stars
Rust 1 awesome list

NVIDIA/Model-Optimizer

A unified library of SOTA model optimization techniques like quantization, distillation, pruning, neural architecture search, speculative decoding, etc. It compresses deep learning models for downstream deployment frameworks like TensorRT-LLM, TensorRT, vLLM, etc. to optimize inference speed.

3,780 stars
Python 2 awesome lists

llm-d/llm-d

Achieve state of the art inference performance with modern accelerators on Kubernetes

4,473 stars
Shell 2 awesome lists