llm-router

llm-router

zxuhan

Cache-aware router for OpenAI-compatible LLM servers, in Go. Per-worker radix trees route each request to the worker holding its KV prefix. Validated on 4x A100 + vLLM and Apple Silicon + llama.cpp.

3 Stars
0 Forks
0 Watchers
Go Language
mit License
42.1 SrcLog Score
Cost to Build
$76.2K
Market Value
$53.8K

Growth over time

2 data points  ·  2026-07-26 → 2026-08-10
Stars Forks Watchers
💬

How do you feel about this project?

Ask AI about llm-router

Question copied to clipboard

What is the zxuhan/llm-router GitHub project? Description: "Cache-aware router for OpenAI-compatible LLM servers, in Go. Per-worker radix trees route each request to the worker holding its KV prefix. Validated on 4x A100 + vLLM and Apple Silicon + llama.cpp.". Written in Go. Explain what it does, its main use cases, key features, and who would benefit from using it.

Question is copied to clipboard — paste it after the AI opens.

How to clone llm-router

Clone via HTTPS

git clone https://github.com/zxuhan/llm-router.git

Clone via SSH

[email protected]:zxuhan/llm-router.git

Download ZIP

Download main.zip

Found an issue?

Report bugs or request features on the llm-router issue tracker:

Open GitHub Issues