aurelio-labs/semantic-router
Superfast AI decision making and intelligent processing of multi-modal data.
A programmable Mixture-of-Models router for heterogeneous LLM inference
Appears on
Quick read
Latest capture 2026-09-10 10:55
12 paths
Agent instructions and tool configuration found in this repository.
Agent instructions 11
GitHub Copilot
1 observed capture since 2026-09-10. Observed captures are shown by default.
Stars from first capture 0
Observed captures only
All tracked data
Observed snapshots
Observed snapshots
Nearest indexed repositories by embedding similarity.
Superfast AI decision making and intelligent processing of multi-modal data.
A high-throughput and memory-efficient inference and serving engine for LLMs
Achieve state of the art inference performance with modern accelerators on Kubernetes
A framework for serving and evaluating LLM routers - save LLM costs without compromising quality
Democratizing Reinforcement Learning for LLMs
Community maintained hardware plugin for vLLM on Ascend