58f2a22966
Adds a new `pinned` group with swap: false (models coexist in VRAM), exclusive: false (group shares with other groups), persistent: true (never unload). Each member also gets ttl: 0 so the per-model idle-timeout can't drop them either — belt + suspenders. Pair is currently qwen3.5-9b (~6 GB Q4) + qwen3.6-35-a3b (~29 GB Q6). Plus the 128K KV caches, roughly 50-60 GB VRAM resident. Appropriate for an A6000/H100-class card; verify fit after deploy. Committed as a canonical change; push + restart still needed on ana-ml2.