Skip to content

multimodal 61

  • modlens v3.25.0 Verified Web UI 3695

    Plug-in vision for text-only LLMs, powered by the free Antigravity CLI

    dsh plugin add @liustack/modlens
  • dsh-vision-router v2.0.1 Verified Web UI 997

    Eyes for text-only DeepSeek Harness agents: built-in free vision chain (no key) + pixel-level vision tools (Q&A, grounding, crop, pixel diff, colors, OCR, SVG trace, cutout, screenshots). One-command install, no Python, image turns work like ordinary tool

    dsh plugin add dsh-vision-router
  • dsh-image-gen v0.2.0 Verified Web UI 219

    Bring ChatGPT-like image generation to DeepSeek Harness — Gemini, OpenAI, Seedream, DashScope & more.

    dsh plugin add dsh-image-gen
  • dsh-crew v0.1.0-rc.6 Verified Web UI 115

    DeepSeek Harness plugin: dispatch work to DSH agents from Claude Code / Codex, as native subagents with live progress

    dsh plugin add @zseven-w/dsh-crew
  • dsh-video-lens v0.3.1 Verified 48

    Give text-only DeepSeek Harness agents video understanding: scene-aware frame sampling + VLM + optional ASR transcript fused into timeline evidence. / 给纯文本模型的视频理解插件(场景感知抽帧 + VLM + 可选语音转录)

    dsh plugin add dsh-video-lens
  • picturereader v3.2.0 Verified Web UI 34

    Unified image understanding plugin for DeepSeek Harness (DSH). Visual twin adapter for native thumbnails + auto-analysis on any text-only model (incl. pi-ai providers); privacy/smart/strict routing; local tools (scan/OCR×3 engines/crop/palette/compare/bat

    dsh plugin add picturereader
  • dsh-vision-proxy v0.4.1 Verified 15

    DeepSeek brain + automatic image transcription for DeepSeek Harness: a deepseek-vision provider route that transcribes attached images to text before delegating to the text-only DeepSeek adapter. Official deepseek-v4-flash-vision-exp by default (a pure-te

    dsh plugin add dsh-vision-proxy
  • dsh-mmx-bridge v1.0.8 Verified Web UI 10

    MiniMax multimodal bridge for DeepSeek Harness (DSH). One mmx_bridge tool covers describe/image/video/speech/music/cover/search/quota; optional web_search/read_image takeover; built-in client enhancement renders inline players/previews plus a settings-pag

    dsh plugin add dsh-mmx-bridge
  • dsh-vision-web v0.1.0 Verified Web UI 8

    Eyes for text-only DeepSeek on DeepSeek Harness: image understanding via Doubao Web by default (zero cost, no API key), Antigravity IDE quota (flash/pro), Gemini API, or Cockpit proxy — model-invokable vision tool, wrapper adapters, evidence memory, and a

    dsh plugin add dsh-vision-web
  • dsh-deepseek-vision v0.1.7 Verified Web UI 8

    Out-of-tree dsh provider plugin: a DeepSeek gateway route that claims image input and transparently describes pasted images through a configured vision-language model (e.g. Qwen-VL) before the text-only DeepSeek wire sees them.

    dsh plugin add dsh-deepseek-vision
  • dsh-tool-vision v0.6.4 Verified Web UI 7

    DeepSeek Harness 外置视觉模型插件:inspect_image 把本地图片或 http(s) 图片 URL 发给任意 OpenAI 兼容端点,视觉模型看图的文字回答直接带回对话;附 Web UI 设置栏。

    dsh plugin add dsh-tool-vision
  • deepseekeyes v0.8.1 Verified Web UI 7

    Auditable vision and cross-platform Computer Use runtime for DeepSeek Harness with source-preserving evidence.

    dsh plugin add @dttxorg/deepseekeyes
  • dsh-vision-bridge v1.0.0 Verified Web UI 6

    On-demand vision for text-only DSH sessions: images become markers, and a vision_describe tool sends only image + question to an OpenAI-compatible vision model

    dsh plugin add @dsh-extension/dsh-vision-bridge
  • dsh-mineru v0.1.9 Verified Web UI 4

    dsh-mineru: MinerU 文档解析插件 for DeepSeek Harness — 多模态全格式 (PDF/Word/PPT/Excel/HTML/图片) → 结构化 Markdown. 填 Token 走精准解析 API, 留空走 Agent 轻量解析 API.

    dsh plugin add dsh-mineru
  • dsh-sight v0.3.1 Verified Web UI 3

    Plug-in vision for text-only DeepSeek Harness (dsh) models: a `vision` tool with built-in cheap/free VLM presets, multi-image batch analysis, paste-to-hint image admission, and a web settings page with hot-reload.

    dsh plugin add dsh-sight
  • dsh-model-capability v1.0.1 Verified Web UI 3

    Explicit model input-modality selection for the DeepSeek Harness Web UI

    dsh plugin add dsh-model-capability
  • dsh-multi-model-provider v0.1.0-rc.11 Verified Web UI 3

    Register models, assist with portraits, and select the Agent model from a secret-free catalog for DeepSeek Harness.

    dsh plugin add dsh-multi-model-provider
  • dsh-multimodal-bridge v0.1.2 Verified 3

    DeepSeek Harness plugin bundle: qwen_vision (Qwen-VL image understanding) and qwen_generate (Qwen-Image text-to-image and image editing) tools for text-only models

    dsh plugin add dsh-multimodal-bridge
  • dsh-open-eyes v0.1.1-rc.2 Verified Web UI 3

    Open Eyes for DeepSeek Harness: delegate images to a configurable multimodal model through OpenAI Responses, Chat Completions, or Anthropic Messages.

    dsh plugin add dsh-open-eyes
  • dsh-vision-mix v0.2.2 Verified Web UI 3

    A DeepSeek Harness Mix plugin for vision routing plus GPT Image generation and editing.

    dsh plugin add dsh-vision-mix
  • dsh-youreyes v0.1.0 Verified Web UI 3

    Eyes for text-only DeepSeek on DeepSeek Harness: image understanding via Antigravity IDE quota (default, flash/pro), OpenAI-compatible VLM endpoints, Gemini API, or local Ollama — model-invokable vision tool, wrapper adapters for deepseek/opencode-go, evi

    dsh plugin add dsh-youreyes
  • dsh-auto-vision v0.4.0 Verified 2

    DeepSeek Harness 视觉桥:自动发现你已配置的多模态模型,给纯文本主模型装上 vision 工具,识别结果以纯文本返回。零配置,一条命令安装。

    dsh plugin add dsh-auto-vision
  • dsh-auto-image v0.1.1 Verified 2

    Bridges images into text for DeepSeek Harness: when the session's selected model cannot see images, a configured vision model describes them and the descriptions enter the durable session history as folded context rows — your message stays untouched.

    dsh plugin add dsh-auto-image
  • dsh-vision-plugin v1.0.4 Verified Web UI 2

    为 DeepSeek Harness 提供外部视觉模型能力:纯文本主模型通过 describe_image 工具调用外部视觉模型看图,拿到纯文字描述(多模态补全)。静态 Cordis 插件,随 DSH 启动自动加载。

    dsh plugin add @woyeshishen/dsh-vision-plugin
  • dsh-vision-recognizer v0.2.0 Verified Web UI 2

    Adaptive image routing for DeepSeek Harness: pass images directly to native multimodal models and transcribe them only for text-only models.

    dsh plugin add dsh-vision-recognizer
  • dsh-mindseye v0.2.7 Verified Web UI 2

    MindsEye: model-driven vision tools, structured evidence, and exact cache for DeepSeek Harness

    dsh plugin add dsh-mindseye
  • dsh-plugin-qwen-image v0.3.1 Verified Web UI 2

    DeepSeek Harness out-of-tree plugin: give a text-only coding model eyes by routing images to a Qwen-VL (DashScope) route through ctx.llm and returning text.

    dsh plugin add dsh-plugin-qwen-image
  • deepsee v4.0.2 Verified Web UI 2

    DeepSee: DeepSeek Harness vision, multi-model routing, and Gemini visual self-checks before delivery

    dsh plugin add @chang416/deepsee
  • dsh-design-qa v0.1.4 Verified 2

    让 DeepSeek Harness 里任何纯文本模型都能读图。识图是按需调用的 tool —— 图片不进主模型上下文,不看就不花钱;附 23 处缺陷的评测集,换模型可自测。

    dsh plugin add dsh-design-qa
  • dsh-autovision v0.1.0 Verified Web UI 1

    dsh auto-vision bridge: paste an image into a text-only model's composer and a configured multimodal model transcribes it as text, automatically.

    dsh plugin add @iroam2375/dsh-autovision
  • dsh-sight v0.1.11 Verified Web UI 1

    DeepSeek Harness plugin: direct multimodal image transfer declarations + per-session image clearing (surface replace), with a settings page and composer controls.

    dsh plugin add @eric.wen/dsh-sight
  • dsh-vision-bridge v0.1.1 Verified 1

    DSH 双模型路由插件:主对话留在 DeepSeek,图片任务自动分发到显式配置的视觉模型,再把分析结果交回 DeepSeek 继续。

    dsh plugin add @acc1143/dsh-vision-bridge
  • dsh-crew v0.5.0 Verified Web UI 1

    DeepSeek Harness plugin: dispatch work to DSH agents from Claude Code / Codex, as native subagents with live progress

    dsh plugin add @ran-sh/dsh-crew
  • dsh-gemini-multimodal v0.1.2 Verified 1

    DeepSeek Harness plugin: multimodal tools (image/audio/video/document understanding, transcription, image generation) via Gemini API or the local Antigravity CLI.

    dsh plugin add dsh-gemini-multimodal
  • dsh-glm-vision v1.0.1 Verified Web UI 1

    GLM 视觉模型插件:注册 glm-vision 供应商路由(glm-4.6v 系列,image+text 输入声明)并提供 glm_vision 工具,让 DeepSeek 等文本主模型直接调用智谱视觉模型看图。

    dsh plugin add dsh-glm-vision
  • dsh-image-plugins v0.1.4 Verified 1

    Multimodal plugin for DeepSeek Harness: understand images and generate images through configurable OpenAI-compatible or DashScope endpoints.

    dsh plugin add dsh-image-plugins
  • dsh-local-vision v0.2.1 Verified 1

    DeepSeek Harness plugin: a model-callable local_vision tool that describes images with a local vision model through Ollama. No cloud, the image never leaves the machine. Two tiers: fast (text/colors/summary) and detailed (full description).

    dsh plugin add dsh-local-vision
  • dsh-mingmu v0.1.5 Verified Web UI 1

    明眸 VisionBridge - 自研视觉桥:瞎子模型收图时自动调用视觉模型识别,把识别文字喂回主模型,无感、可配置、升级不丢

    dsh plugin add dsh-mingmu
  • dsh-plugin-bilibili v0.2.1 Verified 1

    DeepSeek Harness plugin: Bilibili keyword video search, video metadata, subtitle transcripts, direct play URLs, and multimodal frame viewing (bilibili_search / bilibili_video / bilibili_subtitles / bilibili_playurl / bilibili_frames). Anonymous by default

    dsh plugin add dsh-plugin-bilibili
  • dsh-plugin-vision-toolkit v0.1.1 Verified 1

    Vision toolkit for DeepSeek Harness -- glance, ground, detect, crop CLI tools for text-only agents to understand images

    dsh plugin add dsh-plugin-vision-toolkit
  • dsh-tool-visual-primitives v1.3.0 Verified Web UI 1

    DSH 视觉原语工具:参考 DeepSeek《Thinking with Visual Primitives》论文,将图片路由到外部视觉模型并返回带视觉基元的文本分析。纯文本循环,对话模型无需原生视觉能力即可'看见'图片。

    dsh plugin add dsh-tool-visual-primitives
  • dsh-vision-guard v0.1.3 Verified 1

    Transparent image guard + vision analysis for DeepSeek Harness: text-only models read pasted images without the 400 session deadlock.

    dsh plugin add dsh-vision-guard
  • dsh-vision-link v1.2.2 Verified Web UI 1

    Lightweight, route-preserving vision link for DSH: a configured vision model sees while the selected text model stays in control.

    dsh plugin add dsh-vision-link
  • dsh-vision-subagent v0.3.0 Verified Web UI 1

    Vision for text-only DeepSeek Harness agents: delegate image reading to a one-shot subagent on a configurable vision route (MiniMax / Kimi / any OpenAI-compatible provider), keeping image bytes and the vision model's context out of the main session.

    dsh plugin add dsh-vision-subagent
  • soyo-dsh-plugin v0.1.1 Verified 1

    DSH-native video understanding with configurable multimodal providers

    dsh plugin add soyo-dsh-plugin
  • dsh-eyes v0.4.0 Verified Web UI

    A vision bridge profile bundle for DeepSeek Harness (dsh): gives non-vision models image-reading capability by delegating transcription to a vision-capable model, with GUI paste admission and a web settings page for vision model selection.

    dsh plugin add dsh-eyes
  • dsh-vision-adapter v0.6.4 Verified Web UI

    给 DeepSeek Harness 加视觉能力:可视化设置页选视觉厂商(Kimi/智谱/通义/OpenAI/Gemini/豆包/MiniMax/阶跃星辰)并粘贴 API Key,聊天里即可拖图识别——图片走视觉模型、文字走 DeepSeek 推理。

    dsh plugin add dsh-vision-adapter
  • dsh-tool-vision v0.1.2 Verified Web UI

    Vision tools (analyze_image, locate_element) for DeepSeek Harness: local ollama, DSH subagent (qwen-vl etc), or OpenAI-compatible HTTP endpoints

    dsh plugin add @bujue3184/dsh-tool-vision
  • dsh-vision-helper v0.4.2 Verified Web UI

    Persistent vision plugin for DeepSeek Harness: registers the vision_analyze tool backed by a configurable multimodal model, with a settings-page UI. Zero dependencies.

    dsh plugin add dsh-vision-helper
  • dsh-vision-bridge v0.1.7 Verified Web UI

    DSH 视觉桥插件:自动区分多模态/文本模型。多模态模型直接看图;文本模型通过可配置的多模态端点(baseUrl + apiKey + model)代看,支持粘贴图片、read_image 工具、请求时图片转证据。

    dsh plugin add @omdp/dsh-vision-bridge