53f00b232b
AtlaAI's Selene-1-Mini judge model for evaluation/scoring tasks. Llama 3.1 8B base, mradermacher imatrix-quantized Q6_K (~6.5GB, quality-leaning quant). Apache-2.0. Per Atla cookbook these defaults hit 84% on RAGTruth hallucination eval. New 'JUDGE / EVAL MODELS' section between the dense chat models and the embedding models — separate category from chat/reasoning since the run-params shape is different (deterministic-leaning: temp 0.01, top-p 1.0, no repeat penalty). q8_0 KV cache to fit 32K ctx cleanly on the 3090 with headroom. Pre-pulled into the shared HF cache via the new playbooks/pull-hf-model.yaml playbook (canonical replacement for ad-hoc huggingface_hub.snapshot_download calls; see CHANGELOG). Smoke-tested 2026-05-13: GET /v1/models lists selene-1-mini-8b, POST /v1/chat/completions returns expected output cleanly.