llama-vulkan-strix

llama-vulkan-strix

hec-ovi

llama.cpp OpenAI-compatible server on Vulkan for AMD Strix Halo (gfx1151), GGUF weights pinned to GTT not VRAM. Serves poolside Laguna S 2.1, Gemma 4 and Qwen3.6 GGUFs on the stock Vulkan image, plus an opt-in ROCmFP4 + MTP stack (Ubuntu 26.04 + TheRock ROCm 7.13). Docker Compose, with real measured benchmarks.

18 Stars
1 Forks
0 Watchers
Python Language
other License
87.2 SrcLog Score
Cost to Build
$2.6K
Market Value
$4.4K

Growth over time

1 data points  ·  2026-07-26 → 2026-07-26
Stars Forks Watchers
💬

How do you feel about this project?

Ask AI about llama-vulkan-strix

Question copied to clipboard

What is the hec-ovi/llama-vulkan-strix GitHub project? Description: "llama.cpp OpenAI-compatible server on Vulkan for AMD Strix Halo (gfx1151), GGUF weights pinned to GTT not VRAM. Serves poolside Laguna S 2.1, Gemma 4 and Qwen3.6 GGUFs on the stock Vulkan image, plus an opt-in ROCmFP4 + MTP stack (Ubuntu 26.04 + TheRock ROCm 7.13). Docker Compose, with real measured benchmarks.". Written in Python. Explain what it does, its main use cases, key features, and who would benefit from using it.

Question is copied to clipboard — paste it after the AI opens.

How to clone llama-vulkan-strix

Clone via HTTPS

git clone https://github.com/hec-ovi/llama-vulkan-strix.git

Clone via SSH

[email protected]:hec-ovi/llama-vulkan-strix.git

Download ZIP

Download main.zip

Found an issue?

Report bugs or request features on the llama-vulkan-strix issue tracker:

Open GitHub Issues