For brokkr's dataset foundry (operator-approved 2026-09-26, relayed). - stacks/lfm-vl-seat: llama.cpp server-cuda b11176 (digest-pinned), Q5_K_M + mmproj Q8_0, :8030; gateway alias lfm25-vl-3b (LiteLLM restarted, 36 s). Positive control exact; null control shows it describes a missing image. - stacks/vibevoice-asr-seat: audio.cpp v0.8.2-audio8-perf-hotfix (the GGUF's own runtime, not vibevoice.cpp) on cuda 12.8 runtime + libgomp + libsoxr, sha256-pinned; :8031 direct. LibriSpeech WER 3/69, RTF 0.07-0.14; ~31 s cold first request.
16 lines
306 B
JSON
16 lines
306 B
JSON
{
|
|
"host": "0.0.0.0",
|
|
"port": 8080,
|
|
"backend": "cuda",
|
|
"threads": 6,
|
|
"models": [
|
|
{
|
|
"id": "vibevoice-asr-streaming-1.5b",
|
|
"family": "vibevoice_asr_streaming",
|
|
"path": "/models/vibevoice-asr-streaming-1.5b-q4_k.gguf",
|
|
"task": "asr",
|
|
"mode": "streaming"
|
|
}
|
|
]
|
|
}
|