docs: correct sec-seat lineage — M.O.G.-SEC/mog-sec is an offense+defense SFT finetune, not a persona-on-stock

The sentinel-r3 header and two memory notes described mog-sec (Blackfrost
M.O.G.-SEC / Qwentium) as 'a persona system prompt on stock weights'. Its card
is explicit that it is NOT: base_model_relation: finetune on Qwen/Qwen3.8-27B,
a refusal-free offense+defense cybersecurity SFT with YaRN 1M context ('not a
system-prompt sticker on a stock Qwen'). So all three sec-seat candidates are
Qwen3.8-27B SFT finetunes and differ in training focus, not in kind:
mog-sec = broad offense+defense SFT; sentinel-r3 = pentest agent-trajectory SFT;
cyberprev = cyber tool-calling LoRA SFT on an abliterated base.
This commit is contained in:
vh
2026-09-14 07:31:23 -07:00
parent 196416f3da
commit 3906c6842c
3 changed files with 8 additions and 4 deletions
+4 -2
View File
@@ -2,8 +2,10 @@
#
# glyphsoftware/sentinel-r3: a REAL SFT finetune of stock Qwen/Qwen3.8-27B on 1,230
# authorized-pentest agent trajectories (recon -> foothold -> privesc -> writeup) over a
# 19-tool surface that matches our own harness. Contrast mog-sec, which is a persona
# system prompt on stock weights. Quantized in-house to the same mixed NVFP4 W4A4(MLP
# 19-tool surface that matches our own harness. Contrast mog-sec (M.O.G.-SEC/Qwentium),
# which is a refusal-free OFFENSE+DEFENSE cyber SFT of Qwen3.8-27B — its card explicitly says
# "not a system-prompt sticker on a stock Qwen" (an earlier note here called it that; wrong).
# So all three sec seats are Qwen3.8-27B finetunes, differing in training focus. Quantized in-house to the same mixed NVFP4 W4A4(MLP
# 0-55) + FP8 W8A8(attn/lm_head/MLP 56-63) recipe as mog-sec/gen. → sentinel-r3-nvfp4-mixed.PROVENANCE.txt
#
# ⚠ PROPRIETARY LICENSE (Glyph Proprietary v1.0) — operator's fair-use/licensee call, unlike