跳到主要内容

llama-server 3

  • dsh-local-models v0.4.0 已验证 Web 界面

    dsh addon: a Local Models settings tab that starts and stops a llama-server child process living with the dsh host process.

    dsh plugin add dsh-local-models
  • DSH plugin: inject per-model sampling parameters (top_p, top_k, min_p, repeat_penalty, presence_penalty, frequency_penalty) into every chat-completions request sent to a local llama.cpp / llama-server gateway.

    dsh plugin add dsh-llama-cpp-sampling-params
  • dsh-llm-sampling-params v0.2.1 已验证

    DSH plugin: inject per-model sampling parameters (top_p, top_k, min_p, repeat_penalty/repetition_penalty, presence_penalty, frequency_penalty) into every chat-completions request sent to an OpenAI-compatible gateway (llama.cpp, SGLang, vLLM).

    dsh plugin add dsh-llm-sampling-params