Bundle
dsh-local-llm
DeepSeek Harness plugin for local GGUF models without Ollama
wertyBSd2
3 results
DeepSeek Harness plugin for local GGUF models without Ollama
DSH plugin: manage local model inference servers (llama.cpp / vLLM / SGLang) — registry, safe start/stop, parameter profile versioning with validation. Red lines: never pkill, never stop protected port 11437, never auto-restart dsh.
DSH plugin: manages the Windows llama-server.exe model lifecycle (start / stop / switch / recover) behind a stable OpenAI-compatible gateway, with graceful Ctrl+C shutdown that actually frees VRAM.