straylight: local Whisper ASR and Qwen3.8-Flash-Next uncensored
whisper.cpp large-v3-turbo on :11435 (CPU, OpenAI transcriptions path) so ASR does not take GTT from llama-server. Flash-Next is the cygnal IQ4_XS-NGQ4 GGUF (~98 GB, gfx1151), qwen4exp, mmproj pinned; unload Laguna before loading.
This commit is contained in:
@@ -35,6 +35,7 @@ let
|
||||
"ornith-1.5-9b-uncensored" = vision "Ornith 1.5 9B";
|
||||
"ornith-1.5-35b-a3b" = text "Ornith 1.5 35B MoE";
|
||||
"qwen3.8-27b-uncensored" = text "Qwen3.8 27B";
|
||||
"qwen3.8-flash-next-uncensored" = vision "Qwen3.8 Flash Next";
|
||||
"qwen3-vl-8b-abliterated" = vision "Qwen3-VL 8B";
|
||||
};
|
||||
};
|
||||
|
||||
Reference in New Issue
Block a user