ollama-herd

ollama-herd

geeks-accelerator

Local AI load balancer for Ollama and MLX fleets. Auto-discovery, smart routing, OpenAI-compatible API, zero config. Perfect for Mac Minis & Studios.

17 Stars
1 Forks
0 Watchers
Python Language
mit License
81.2 SrcLog Score
Cost to Build
$99.3K
Market Value
$162.7K

Growth over time

5 data points  ·  2026-04-11 → 2026-08-10
Stars Forks Watchers
💬

How do you feel about this project?

Ask AI about ollama-herd

Question copied to clipboard

What is the geeks-accelerator/ollama-herd GitHub project? Description: "Local AI load balancer for Ollama and MLX fleets. Auto-discovery, smart routing, OpenAI-compatible API, zero config. Perfect for Mac Minis & Studios.". Written in Python. Explain what it does, its main use cases, key features, and who would benefit from using it.

Question is copied to clipboard — paste it after the AI opens.

How to clone ollama-herd

Clone via HTTPS

git clone https://github.com/geeks-accelerator/ollama-herd.git

Clone via SSH

[email protected]:geeks-accelerator/ollama-herd.git

Download ZIP

Download main.zip

Found an issue?

Report bugs or request features on the ollama-herd issue tracker:

Open GitHub Issues