refactor(playbooks): host-generic GPU host + GPU LXC playbooks for nh3-ml1

- esh-pve-nvidia-host -> pve-nvidia-host: headers/dkms/build-essential step,
  nouveau blacklist + guarded unload (refuses if nouveau bound a device)
- esh-ml1-lxc -> gpu-lxc: host vars have no defaults (elway aborts on undefined),
  rootfs storage/startup order parameterized, CT kept out of all-guests vzdump jobs
- embed-rerank: Homepage labels take HOST_NAME/HOST_IP, defaults = esh-ml1
This commit is contained in:
vh
2026-09-25 14:17:31 -07:00
parent cdd7605e89
commit bc278d4ba8
10 changed files with 170 additions and 72 deletions
+1 -1
View File
@@ -167,7 +167,7 @@ Pre-flight done before shutdown:
8. **pfi-gx10:** if Prime ran the AC-pull test on this visit, verify it came back
by itself.
9. **Next work:** the GPU's purpose is TBD with Prime; likely the second embed/rerank
backend. Reuse `playbooks/esh-pve-nvidia-host.yaml` / `esh-ml1-lxc.yaml` (make
backend. Reuse `playbooks/pve-nvidia-host.yaml` / `gpu-lxc.yaml` (make
them host-generic). nh3-pve runs kernel 6.8.12-11 with **no matching
proxmox-headers installed**, on PVE 8.4.1 (Debian 12 template only). If the LM
port was cabled, MEBx provisioning can be done through the NanoKVM.