The control plane for self-hosted AI inference. Warm-state GPU routing, multi-runtime orchestration across Ollama, vLLM, llama.cpp, TGI and MLX . Single Go binary. Apache-2.0.
Automated Zabbix Agent deployment and registration pipeline for Hybrid Cloud (Windows/Linux) using Ansible, WinRM, and REST APIs.