mirror of
https://github.com/mudler/LocalAI.git
synced 2026-09-28 17:15:02 -04:00
fix(compose): request NVIDIA compute capability (#11990)
The legacy NVIDIA device reservation requests utility without compute. Docker derives driver capabilities from that list, leaving CUDA libraries unavailable even when monitoring works. Include compute in the legacy example and clarify the matching docs. Assisted-by: Codex:GPT-6 Co-authored-by: localai-org-maint-bot <306269227+localai-org-maint-bot@users.noreply.github.com>
This commit is contained in:
1 parent
c6f1e96a7d
commit
1b1bd0f069
3 files changed
+12
-5
No files matched your search
@@ -88,8 +88,10 @@ page in the frontend shows the node as fully used, check two things:
|
||||
NVML work inside the container. With `--gpus all` alone (or
|
||||
`--runtime nvidia` without extra flags) only `compute` is wired in on
|
||||
some driver versions. Add `-e NVIDIA_DRIVER_CAPABILITIES=compute,utility`
|
||||
to your `docker run`, or `capabilities: [gpu, utility]` in compose /
|
||||
Kubernetes device reservations.
|
||||
to your `docker run`. For Docker Compose with `driver: nvidia`, use
|
||||
`capabilities: [gpu, compute, utility]` on the device reservation.
|
||||
Include `compute` for CUDA libraries such as `libcuda.so.1`; `utility`
|
||||
alone only provides monitoring libraries and tools.
|
||||
2. Pass `--init` to `docker run` (or `init: true` in compose) so the
|
||||
container has a proper PID 1 reaper - otherwise short-lived child
|
||||
processes like `nvidia-smi` can intermittently fail with
|
||||
|
||||
Reference in new issue
Block a user