straylight: local Whisper ASR and Qwen3.8-Flash-Next uncensored

whisper.cpp large-v3-turbo on :11435 (CPU, OpenAI transcriptions path)
so ASR does not take GTT from llama-server. Flash-Next is the cygnal
IQ4_XS-NGQ4 GGUF (~98 GB, gfx1151), qwen4exp, mmproj pinned; unload
Laguna before loading.
This commit is contained in:
2026-09-16 07:18:28 -07:00
parent 927eff0c98
commit 1e94e6638e
2 changed files with 75 additions and 1 deletions
+1
View File
@@ -35,6 +35,7 @@ let
"ornith-1.5-9b-uncensored" = vision "Ornith 1.5 9B";
"ornith-1.5-35b-a3b" = text "Ornith 1.5 35B MoE";
"qwen3.8-27b-uncensored" = text "Qwen3.8 27B";
"qwen3.8-flash-next-uncensored" = vision "Qwen3.8 Flash Next";
"qwen3-vl-8b-abliterated" = vision "Qwen3-VL 8B";
};
};