eth-sri/lmql
A language for constraint-guided and efficient LLM programming.
Run LLMs with MLX
Appears on
Quick read
Latest capture 2026-08-25 03:04
9 observed captures since 2026-05-22. Observed captures are shown by default.
Stars from first capture +1369
Observed captures only
All tracked data
Observed snapshots
Observed snapshots
Nearest indexed repositories by embedding similarity.
A language for constraint-guided and efficient LLM programming.
llama.cpp fork with additional SOTA quants and improved performance
Running large language models on a single GPU for throughput-oriented scenarios.
A fast inference library for running LLMs locally on modern consumer-class GPUs
Fast Multimodal LLM on Mobile Devices
A framework for few-shot evaluation of language models.