AMD ROCm GPU inference for Linux. llama.cpp + HIP backend for RDNA 3.
Find a file
Snider aa42cff417 feat: scaffold go-rocm AMD GPU inference package
Implements inference.Backend via llama-server subprocess (llama.cpp + HIP/ROCm).
Targets RX 7800 XT (gfx1101, RDNA 3, 16GB VRAM).

Includes:
- Backend registration with build tags (linux/amd64)
- Stub backend.go with llama-server lifecycle outline
- CLAUDE.md with build instructions for llama.cpp + ROCm
- TODO.md with 5-phase task queue
- FINDINGS.md with hardware specs, VRAM budget, design rationale

Co-Authored-By: Virgil <virgil@lethean.io>
2026-02-19 19:39:40 +00:00
backend.go feat: scaffold go-rocm AMD GPU inference package 2026-02-19 19:39:40 +00:00
CLAUDE.md feat: scaffold go-rocm AMD GPU inference package 2026-02-19 19:39:40 +00:00
FINDINGS.md feat: scaffold go-rocm AMD GPU inference package 2026-02-19 19:39:40 +00:00
go.mod feat: scaffold go-rocm AMD GPU inference package 2026-02-19 19:39:40 +00:00
README.md Initial commit 2026-02-19 19:35:55 +00:00
register_rocm.go feat: scaffold go-rocm AMD GPU inference package 2026-02-19 19:39:40 +00:00
rocm.go feat: scaffold go-rocm AMD GPU inference package 2026-02-19 19:39:40 +00:00
rocm_stub.go feat: scaffold go-rocm AMD GPU inference package 2026-02-19 19:39:40 +00:00
TODO.md feat: scaffold go-rocm AMD GPU inference package 2026-02-19 19:39:40 +00:00

go-rocm

AMD ROCm GPU inference for Linux. llama.cpp + HIP backend for RDNA 3.