multimodal 75
Plug-in vision for text-only LLMs, powered by the free Antigravity CLI
dsh plugin add @liustack/modlensEyes for text-only DeepSeek Harness agents: built-in free vision chain + pixel-level tools. Remote vision providers, including the default free fallback, receive image content unless strict local-only use is configured.
dsh plugin add dsh-vision-routerBring ChatGPT-like image generation to DeepSeek Harness — Gemini, OpenAI, Seedream, DashScope, local ComfyUI & more.
dsh plugin add dsh-image-genDeepSeek Harness plugin: dispatch work to DSH agents from Claude Code / Codex, as native subagents with live progress
dsh plugin add @zseven-w/dsh-crewGive text-only DeepSeek Harness agents video understanding: scene-aware frame sampling + VLM + optional ASR transcript fused into timeline evidence. / 给纯文本模型的视频理解插件(场景感知抽帧 + VLM + 可选语音转录)
dsh plugin add dsh-video-lens让 DeepSeek Harness 里任何纯文本模型都能读图。识图是按需调用的 tool —— 图片不进主模型上下文,不看就不花钱;附 23 处缺陷的评测集,换模型可自测。
dsh plugin add dsh-design-qaUnified image understanding plugin for DeepSeek Harness (DSH). Visual twin adapter for native thumbnails + auto-analysis on any text-only model (incl. pi-ai providers); privacy/smart/strict routing; local tools (scan/OCR×4 engines (windows/macos/paddle/ra
dsh plugin add picturereaderDeepSeek brain + automatic image transcription for DeepSeek Harness: a deepseek-vision provider route that transcribes attached images to text before delegating to the text-only DeepSeek adapter. Official deepseek-v4-flash-vision-exp by default (a pure-te
dsh plugin add dsh-vision-proxyPer-model reasoning efforts and input modalities for DeepSeek Harness — one settings section, no YAML.
dsh plugin add better-model-providerMiniMax multimodal bridge for DeepSeek Harness (DSH). One mmx_bridge tool covers describe/image/video/speech/music/cover/search/quota; optional web_search/read_image takeover; built-in client enhancement renders inline players/previews plus a settings-pag
dsh plugin add dsh-mmx-bridgeOut-of-tree dsh provider plugin: a DeepSeek gateway route that claims image input and transparently describes pasted images through a configured vision-language model (e.g. Qwen-VL) before the text-only DeepSeek wire sees them.
dsh plugin add dsh-deepseek-visionEyes for text-only DeepSeek on DeepSeek Harness: image understanding via Doubao Web by default (zero cost, no API key), Antigravity IDE quota (flash/pro), Gemini API, or Cockpit proxy — model-invokable vision tool, wrapper adapters, evidence memory, and a
dsh plugin add dsh-vision-webDeepSeek Harness 外置视觉模型插件:inspect_image 把本地图片或 http(s) 图片 URL 发给任意 OpenAI 兼容端点,视觉模型看图的文字回答直接带回对话;附 Web UI 设置栏。
dsh plugin add dsh-tool-visionAuditable vision and cross-platform Computer Use runtime for DeepSeek Harness with source-preserving evidence.
dsh plugin add @dttxorg/deepseekeyesOn-demand vision for text-only DSH sessions: images become markers, and a vision_describe tool sends only image + question to an OpenAI-compatible vision model
dsh plugin add @dsh-extension/dsh-vision-bridgedsh-mineru: MinerU 文档解析插件 for DeepSeek Harness — 多模态全格式 (PDF/Word/PPT/Excel/HTML/图片) → 结构化 Markdown. 填 Token 走精准解析 API, 留空走 Agent 轻量解析 API.
dsh plugin add dsh-mineruDeepSeek vision bridge for the dsh web GUI: route image attachments to a vision model (pi-ai / llama.cpp Qwen3-VL) and continue the conversation with the text description on a text-only LLM (DeepSeek)
dsh plugin add dsh-llm-vision-bridgeDeepSeek Harness 全能插件:识别/生图/改图一体化,无需切换模型——用常规 DeepSeek 模型即可自动调用视觉与生图模型(gemini_vision / gemini_generate_image / gemini_optimize_image)。多后端:Gemini 原生 + 任意 OpenAI 兼容服务(GPT-4o、Qwen-VL、GLM-4V、Moonshot、gpt-image、DALL-E、Flux、Stable Diffusion、OpenRouter、硅基流动、各类中转
dsh plugin add dsh-vision-imagenA DeepSeek Harness Mix plugin for vision routing plus GPT Image generation and editing.
dsh plugin add dsh-vision-mixPlug-in vision for text-only DeepSeek Harness (dsh) models: a `vision` tool with built-in cheap/free VLM presets, multi-image batch analysis, paste-to-hint image admission, and a web settings page with hot-reload.
dsh plugin add dsh-sightExplicit model input-modality selection for the DeepSeek Harness Web UI
dsh plugin add dsh-model-capabilityDeepSeek Harness plugin bundle: qwen_vision (Qwen-VL image understanding) and qwen_generate (Qwen-Image text-to-image and image editing) tools for text-only models
dsh plugin add dsh-multimodal-bridgeDeepSeek Harness plugin: Bilibili keyword video search, video metadata, subtitle transcripts, direct play URLs, and multimodal frame viewing (bilibili_search / bilibili_video / bilibili_subtitles / bilibili_playurl / bilibili_frames). Anonymous by default
dsh plugin add dsh-plugin-bilibiliEyes for text-only DeepSeek on DeepSeek Harness: image understanding via Antigravity IDE quota (default, flash/pro), OpenAI-compatible VLM endpoints, Gemini API, or local Ollama — model-invokable vision tool, wrapper adapters for deepseek/opencode-go, evi
dsh plugin add dsh-youreyesDeepSeek Harness out-of-tree plugin: give a text-only coding model eyes by routing images to a Qwen-VL (DashScope) route through ctx.llm and returning text.
dsh plugin add dsh-plugin-qwen-image为 DeepSeek Harness 提供外部视觉模型能力:纯文本主模型通过 describe_image 工具调用外部视觉模型看图,拿到纯文字描述(多模态补全)。静态 Cordis 插件,随 DSH 启动自动加载。
dsh plugin add @woyeshishen/dsh-vision-pluginMultimodal plugin for DeepSeek Harness: understand images and generate images through configurable OpenAI-compatible or DashScope endpoints.
dsh plugin add dsh-image-pluginsFlagship multimodal vision hub for DeepSeek Harness: ~40 tools, PDF drag-and-drop, LaTeX formulas, complex tables, QR codes, UI flow diagrams, and multi-model consensus.
dsh plugin add @goodandready/dsh-vision-bridgeRegister models, assist with portraits, and select the Agent model from a secret-free catalog for DeepSeek Harness.
dsh plugin add dsh-multi-model-providerOpen Eyes for DeepSeek Harness: delegate images to a configurable multimodal model through OpenAI Responses, Chat Completions, or Anthropic Messages.
dsh plugin add dsh-open-eyesAdaptive image routing for DeepSeek Harness: pass images directly to native multimodal models and transcribe them only for text-only models.
dsh plugin add dsh-vision-recognizerDSH 视觉原语工具:参考 DeepSeek《Thinking with Visual Primitives》论文,将图片路由到外部视觉模型并返回带视觉基元的文本分析。纯文本循环,对话模型无需原生视觉能力即可'看见'图片。
dsh plugin add dsh-tool-visual-primitivesDSH 双模型路由插件:主对话留在 DeepSeek,图片任务自动分发到显式配置的视觉模型,再把分析结果交回 DeepSeek 继续。
dsh plugin add @acc1143/dsh-vision-bridgeDeepSee: DeepSeek Harness vision, multi-model routing, and Gemini visual self-checks before delivery
dsh plugin add @chang416/deepseedsh auto-vision bridge: paste an image into a text-only model's composer and a configured multimodal model transcribes it as text, automatically.
dsh plugin add @iroam2375/dsh-autovisionSet image input and reasoning levels for third-party provider models.
dsh plugin add @jcy2387/dsh-models-input-modalitiesDeepSeek Harness plugin: dispatch work to DSH agents from Claude Code / Codex, as native subagents with live progress
dsh plugin add @ran-sh/dsh-crewPoint-and-shoot screenshot capture for DeepSeek Harness: clipboard watcher + system floating window (comment & key-point, copy/save-doc/save-image) + zero-config OCR through the host's own multimodal model (Tongyi Qianwen fallback) + Obsidian per-day merg
dsh plugin add dsh-screenshot-captureVision toolkit for DeepSeek Harness -- glance, ground, detect, crop CLI tools for text-only agents to understand images
dsh plugin add dsh-plugin-vision-toolkitLightweight, route-preserving vision link for DSH: a configured vision model sees while the selected text model stays in control.
dsh plugin add dsh-vision-linkDSH-native video understanding with configurable multimodal providers
dsh plugin add soyo-dsh-pluginDeepSeek Harness 视觉桥:自动发现你已配置的多模态模型,给纯文本主模型装上 vision 工具,识别结果以纯文本返回。零配置,一条命令安装。
dsh plugin add dsh-auto-visionGLM 视觉模型插件:注册 glm-vision 供应商路由(glm-4.6v 系列,image+text 输入声明)并提供 glm_vision 工具,让 DeepSeek 等文本主模型直接调用智谱视觉模型看图。
dsh plugin add dsh-glm-visionMindsEye: model-driven vision tools, structured evidence, and exact cache for DeepSeek Harness
dsh plugin add dsh-mindseyePer-model vision (image input) toggle for DeepSeek Harness (dsh): list every configured model and flip a switch to enable/disable image support without hand-editing settings.yaml. 为 DeepSeek Harness 提供按模型的「支持图片」开关:无需手改 settings.yaml。
dsh plugin add @lijian-ui/dsh-vision-toggleBridges images into text for DeepSeek Harness: when the session's selected model cannot see images, a configured vision model describes them and the descriptions enter the durable session history as folded context rows — your message stays untouched.
dsh plugin add dsh-auto-imageVisionForge: vision understanding + image generation for DeepSeek Harness (DSH)
dsh plugin add @lr611/visionforgeDeepSeek Harness plugin: auto-route image-bearing requests to deepseek-v4-flash-vision-exp, then fall back to the original model.
dsh plugin add dsh-vision-autoswitchCapability-aware vision for DeepSeek Harness: native image input for multimodal models, with an isolated one-shot vision-route fallback for text-only models.
dsh plugin add dsh-vision-subagentDeepSeek Harness plugin: analyse images out of band with a vision model — pasted images are digested into text before admission and a describe_image tool covers image paths, all without ever changing the session's model.
dsh plugin add dsh-image-router