will this work on AMD card

#13
by tapanpatro - opened

Any info on AMD support or vulkan support for this model ?

I built the prismML-Eng/llM.cpp for a RX 6800XT with ROCm 7.2.3 with these flags:

Compiler combo proven on this machine : hipcc as CXX, ROCm LLVM clang as C.
-DGGML_CUDA_FA_ALL_QUANTS=ON is REQUIRED (see Known issues - mixed K/V attention aborts without it,
and ggml-hip compiles the CUDA .cuh sources, so it applies on AMD too).

export ROCM_PATH=/opt/rocm-7.2.3
cmake -S . -B build \
  -DCMAKE_BUILD_TYPE=Release \
  -DGGML_HIP=ON \
  -DGPU_TARGETS=gfx1030 \
  -DGGML_CUDA_FA_ALL_QUANTS=ON \
  -DCMAKE_CXX_COMPILER=/usr/bin/hipcc \
  -DCMAKE_C_COMPILER=/opt/rocm-7.2.3/lib/llvm/bin/clang
cmake --build build --parallel "$(nproc)"

Running their latest HIP/ROCm llama.cpp build currently just fine on a 9070XT, huge context (for this GPU) with this model and I haven't had any issues with running out of context or compaction but it does seem to get stuck for a while after tool calls.

Don't waste your time with ROCM, just use the Vulkan backend. I get 40t/s on R9 390 era cards with it.

When will they release the official version with ROCM or Vulkan support for the 6900XT?

Sign up or log in to comment