fix llamacpp image
This commit is contained in:
@@ -23,7 +23,8 @@ without re-downloading.
|
||||
|
||||
## GPU / Vulkan
|
||||
|
||||
The `server-vulkan` image bundles the Mesa/RADV Vulkan driver, which supports
|
||||
The `server-vulkan` image (`ghcr.io/ggml-org/llama.cpp:server-vulkan`) bundles
|
||||
the Mesa/RADV Vulkan driver, which supports
|
||||
the Radeon 8060S (RDNA 3.5). Full layer offload (`-ngl 999`) puts the ~16 GiB
|
||||
Q4 model entirely in the 96 GiB VRAM pool.
|
||||
|
||||
|
||||
Reference in New Issue
Block a user