Reproducible LLM KV-cache tiered-storage benchmark on AMD MI308X: LMCache parallel-read patch, load clients, orchestration, measured results
What is the mingxin-tech/mingxin-kvcache-bench GitHub project? Description: "Reproducible LLM KV-cache tiered-storage benchmark on AMD MI308X: LMCache parallel-read patch, load clients, orchestration, measured results". Written in Python. Explain what it does, its main use cases, key features, and who would benefit from using it.
Question is copied to clipboard — paste it after the AI opens.
Clone via HTTPS
Clone via SSH
Download ZIP
Download main.zipReport bugs or request features on the mingxin-kvcache-bench issue tracker:
Open GitHub Issues