Skip to content
#

rocmfp4

Here are 2 public repositories matching this topic...

Language: All
Filter by language

llama.cpp OpenAI-compatible server on Vulkan for AMD Strix Halo (gfx1151), GGUF weights pinned to GTT not VRAM. Serves poolside Laguna S 2.1, Gemma 4 and Qwen3.6 GGUFs on the stock Vulkan image, plus an opt-in ROCmFP4 + MTP stack (Ubuntu 26.04 + TheRock ROCm 7.13). Docker Compose, with real measured benchmarks.

  • Updated Jul 24, 2026
  • Python

Improve this page

Add a description, image, and links to the rocmfp4 topic page so that developers can more easily learn about it.

Curate this topic

Add this topic to your repo

To associate your repository with the rocmfp4 topic, visit your repo's landing page and select "manage topics."

Learn more