All writing
Tag

Inference.

1 post ·RSS

Latest
AMD · 04 Sept 2026 · 6 min read

Between a ROCm and a Hard Place.

Picking a llama.cpp backend (Vulkan or ROCm) and a ROCm source (Fedora's packages or AMD's TheRock SDK) on Strix Halo (gfx1151).

Read the article
All writing