2 repositories on SrcLog
Self-hosted LLM/RAG stack in one command — AMD Strix Halo / x86_64 (ROCm/Vulkan, Docker Compose)
Multi-slot LLM inference on AMD Strix Halo: recipes + honest benchmarks (236 tok/s @ 32 streams, llama.cpp Vulkan)