asr 20
Give text-only DeepSeek Harness agents video understanding: scene-aware frame sampling + VLM + optional ASR transcript fused into timeline evidence. / 给纯文本模型的视频理解插件(场景感知抽帧 + VLM + 可选语音转录)
dsh plugin add dsh-video-lensDSH 专属语音输入插件:点按或按住 Alt 说话、松开/再点按转文字,支持热词替换表(hot.txt)、自定义润色提示词、录音电平指示。默认本地离线识别(SenseVoice,零配置零 key、音频不出本机),自动回退浏览器 Web Speech,可选云端 ASR 与润色(复用 DSH 模型)。Voice input for DeepSeek Harness: tap or hold Alt to talk, get text in the composer — local SenseVoice by
dsh plugin add dsh-voice-scribeVoice input plugin for DeepSeek Harness: microphone, transcription, and optional dsh LLM polishing into an editable draft
dsh plugin add dsh-earsFull-duplex voice plugin for DeepSeek Harness: local zipformer2 streaming ASR (no API key) → editable draft; Edge TTS or local VITS / Kokoro read-aloud with live captions; true barge-in; hardened HTTP surface + model SHA256 pinning; compatible with all ds
dsh plugin add dsh-voice-modeLocal recording and meeting transcription for DSH on macOS (Apple Silicon only)
dsh plugin add @huliux/dsh-asr-pluginHold-to-talk voice input for the DeepSeek Harness Web composer: hold the mouse on the input box, speak, see the running transcript float above it, release to insert the text into the draft. Recognition runs locally (SenseVoice via sherpa-onnx) — no API ke
dsh plugin add dsh-hold-to-talkDeepSeek Harness Web GUI plugin for Seminar Copilot: one-click local seminar recording, ASR backend settings (local MLX Whisper / Aliyun Bailian), auto-start of the dev stack.
dsh plugin add seminar-copilot-dsh-pluginLive2D session view with realtime voice: continuous voice input, streaming TTS, subtitle translation, multi-model catalog, standalone entry / Live2D 实时语音会话视图:连续语音输入、流式 TTS、字幕翻译、多模型目录、独立入口
dsh plugin add dsh-live2d-voiceSenseVoiceSmall-powered local speech-to-text for the DSH Desktop input box: click the mic beside the composer, speak, and the recognized draft (with language/emotion tags) is written into the input.
dsh plugin add dshtools-sensevoice-inputContext-aware voice input for DeepSeek Harness with Web Speech, local SenseVoice transcription, model polish, editable Composer drafts, and user-controlled sending
dsh plugin add dsh-dictateDSH voice input plugin: browser realtime + VAD dictation + local/cloud ASR + DeepSeek AI polish
dsh plugin add dsh-plugin-voice-input聊天框语音输入按钮 for DeepSeek Harness: 点击麦克风说话,多引擎转写(智谱 GLM-ASR-2512 / 本地 faster-whisper / Gemini / OpenAI)自动填入输入框。一个按钮,所见即所得。
dsh plugin add dsh-audio-copilotDeepSeek Harness 云端语音输入插件,支持快捷键录音、阿里云百炼实时识别,并将文字追加到当前会话草稿。
dsh plugin add dsh-listenerSenseVoice 语音输入插件 for DeepSeek Harness:在对话输入框旁添加麦克风按钮,录音后调用本地 SenseVoice 服务转成文本填入输入框。首次使用自动下载模型并显示进度,后端由插件自动启动。
dsh plugin add @opensquad/dsh-voice-inputMiMo (Xiaomi) tools as a DSH profile plugin: web search, image/audio/video understanding, ASR transcription, TTS, voice design and voice cloning — native agent tools with a Settings page for the API key.
dsh plugin add @yyfather/dsh-mimo-pluginAnimate any photo into a responsive virtual girl. She talks, turns, smiles, and moves naturally in sync with your conversation. Low-lag, high-detail.
dsh plugin add dsh-live-talkTencent Cloud ASR voice input for DeepSeek Harness: composer mic + live preview + settings card, installed as a pure pluggable bundle (no core changes)
dsh plugin add dsh-tencent-voice-inputVoice input plugin (dual-face): mic button beside the composer send action, live recording bubble streaming PCM to a local Qwen3-ASR python service managed by the host, plus an ASR environment settings page (clone runtime/model repos, create venv, run ser
dsh plugin add dsh-voice-input-qwen-asrFull-duplex voice mode for DeepSeek Harness 0.2 desktop: streaming ASR to draft, sentence read-aloud, barge-in, live captions, plus an in-app settings page. An unofficial compatibility fork of dsh-voice-mode 0.7.15.
dsh plugin add @nutshelllee/dsh-voice-mode-desktop开口即成文 · Speak-to-prompt for DeepSeek Harness:云端 ASR 语音识别 + 提示词优化 + 填入草稿/自动发送,跨平台 macOS / Windows。
dsh plugin add @bittersmilezzz/dsh-asr-voice