strix-halo-quant-lab

strix-halo-quant-lab

kingjones30

Run large LLMs locally on AMD Ryzen AI Max+ 395 (Strix Halo, gfx1151) with ROCmFP4 4-bit quantization. Measured benchmarks, build + serving recipes, and 118 ready-to-run GGUF models.

6 Stars
0 Forks
0 Watchers
Python Language
mit License
63.8 SrcLog Score
Cost to Build
$6.7K
Market Value
$7.5K

Growth over time

2 data points  ·  2026-09-06 → 2026-10-03
Stars Forks Watchers
💬

How do you feel about this project?

Ask AI about strix-halo-quant-lab

Question copied to clipboard

What is the kingjones30/strix-halo-quant-lab GitHub project? Description: "Run large LLMs locally on AMD Ryzen AI Max+ 395 (Strix Halo, gfx1151) with ROCmFP4 4-bit quantization. Measured benchmarks, build + serving recipes, and 118 ready-to-run GGUF models.". Written in Python. Explain what it does, its main use cases, key features, and who would benefit from using it.

Question is copied to clipboard — paste it after the AI opens.

How to clone strix-halo-quant-lab

Clone via HTTPS

git clone https://github.com/kingjones30/strix-halo-quant-lab.git

Clone via SSH

[email protected]:kingjones30/strix-halo-quant-lab.git

Download ZIP

Download main.zip

Found an issue?

Report bugs or request features on the strix-halo-quant-lab issue tracker:

Open GitHub Issues