vision 92
Plug-in vision for text-only LLMs, powered by the free Antigravity CLI
dsh plugin add @liustack/modlensEyes for text-only DeepSeek Harness agents: built-in free vision chain (no key) + pixel-level vision tools (Q&A, grounding, crop, pixel diff, colors, OCR, SVG trace, cutout, screenshots). One-command install, no Python, image turns work like ordinary tool
dsh plugin add dsh-vision-routerUnified image understanding plugin for DeepSeek Harness (DSH). Visual twin adapter for native thumbnails + auto-analysis on any text-only model (incl. pi-ai providers); privacy/smart/strict routing; local tools (scan/OCR×3 engines/crop/palette/compare/bat
dsh plugin add picturereaderDeepSeek Visionary native plugin for DeepSeek Harness: deepseek_vision / status / login / logout native tools plus the text-model image bridge, all backed by the visionary-server CLI (DeepSeek web vision model, no API key).
dsh plugin add @xlight-oss/visionary-dshDeepSeek brain + automatic image transcription for DeepSeek Harness: a deepseek-vision provider route that transcribes attached images to text before delegating to the text-only DeepSeek adapter. Official deepseek-v4-flash-vision-exp by default (a pure-te
dsh plugin add dsh-vision-proxyGenerate and display images inline in the DSH chat via API channels or local CLIs (mmx / codex / agy), and read images into structured JSON evidence (OCR / layout / semantics) on any model — with backend probing and a bundled recovery skill.
dsh plugin add dsh-chat-imagineEyes for text-only DeepSeek on DeepSeek Harness: image understanding via Doubao Web by default (zero cost, no API key), Antigravity IDE quota (flash/pro), Gemini API, or Cockpit proxy — model-invokable vision tool, wrapper adapters, evidence memory, and a
dsh plugin add dsh-vision-webOut-of-tree dsh provider plugin: a DeepSeek gateway route that claims image input and transparently describes pasted images through a configured vision-language model (e.g. Qwen-VL) before the text-only DeepSeek wire sees them.
dsh plugin add dsh-deepseek-visionAuditable vision and cross-platform Computer Use runtime for DeepSeek Harness with source-preserving evidence.
dsh plugin add @dttxorg/deepseekeyesDeepSeek Harness 外置视觉模型插件:inspect_image 把本地图片或 http(s) 图片 URL 发给任意 OpenAI 兼容端点,视觉模型看图的文字回答直接带回对话;附 Web UI 设置栏。
dsh plugin add dsh-tool-visionOn-demand vision for text-only DSH sessions: images become markers, and a vision_describe tool sends only image + question to an OpenAI-compatible vision model
dsh plugin add @dsh-extension/dsh-vision-bridgeDeepSeek Harness plugin that uses configured model providers for image analysis and context compaction.
dsh plugin add @dsh-plugin/dsh-auxiliaryCodex-backed web search, image generation, and image understanding tools for the DeepSeek Harness.
dsh plugin add dsh-codex-toolsThird-party provider reasoning-effort AND input-modality settings for DeepSeek Harness: thinking levels and image-input support declared per model, auto-adapted from a model knowledge base + wire-protocol inference, edited right inside the official Models
dsh plugin add dsh-better-reasoning-effortDeepSeek Harness 原生视觉 Bundle:粘贴或拖入图片,通过托管的 deepseek-vision-mcp 调用 OpenAI 兼容视觉模型。
dsh plugin add dsh-plugin-deepseek-visionDeepSeek Harness (dsh) web plugin: API balance + today's spend sidebar widget, /api/dsh-usage route, external vision-call accounting (JSONL usage log), plus a full per-session usage dashboard (call log, per-model stats, cache rate, CSV export) via the usa
dsh plugin add dsh-usage-dashboard-plusDeepSeek Harness native plugin: vision capabilities for text-only LLMs (describe, OCR, VQA, layout analysis, clipboard)
dsh plugin add dsh-plugin-deepeyePlug-in vision for text-only DeepSeek Harness (dsh) models: a `vision` tool with built-in cheap/free VLM presets, multi-image batch analysis, paste-to-hint image admission, and a web settings page with hot-reload.
dsh plugin add dsh-sightDeepSeek Harness plugin that turns UI screenshots into implementation-grade web specs using OCR, deterministic geometry, scene graphs, assets, and render comparison.
dsh plugin add dsh-ui-specA DeepSeek Harness Mix plugin for vision routing plus GPT Image generation and editing.
dsh plugin add dsh-vision-mixEyes for text-only DeepSeek on DeepSeek Harness: image understanding via Antigravity IDE quota (default, flash/pro), OpenAI-compatible VLM endpoints, Gemini API, or local Ollama — model-invokable vision tool, wrapper adapters for deepseek/opencode-go, evi
dsh plugin add dsh-youreyesDeepSeek Harness plugin: let text-only models receive pasted images, and analyze them with a built-in OpenAI-compatible vision tool
dsh plugin add dsh-image-pathifyDeepSeek Harness plugin bundle: qwen_vision (Qwen-VL image understanding) and qwen_generate (Qwen-Image text-to-image and image editing) tools for text-only models
dsh plugin add dsh-multimodal-bridgeOpen Eyes for DeepSeek Harness: delegate images to a configurable multimodal model through OpenAI Responses, Chat Completions, or Anthropic Messages.
dsh plugin add dsh-open-eyes让 DeepSeek Harness 里任何纯文本模型都能读图。识图是按需调用的 tool —— 图片不进主模型上下文,不看就不花钱;附 23 处缺陷的评测集,换模型可自测。
dsh plugin add dsh-design-qaDeepSeek Harness out-of-tree plugin: give a text-only coding model eyes by routing images to a Qwen-VL (DashScope) route through ctx.llm and returning text.
dsh plugin add dsh-plugin-qwen-imageConnect your sub2api gateway to DeepSeek Harness: OpenAI-compatible multi-provider routes (OpenAI / Claude / Grok / Gemini) behind one base URL, with per-key model discovery, usage lookup, global vision/image tools, and a settings page.
dsh plugin add @godd6366/dsh-sub2apiModel-facing generate_image tool for DeepSeek Harness (dsh): generates brand-new images with the Google Gemini image-generation service via the Antigravity CLI, saves them, and returns the paths. Hot-pluggable — install with `dsh plugin --profile web add
dsh plugin add dsh-tool-generate-imageScreenshot feedback for DeepSeek Harness — let your coding agent SEE what it builds
dsh plugin add dsh-screenshot-feedback-hook-mcpDeepSeek Harness plugin: a model-facing `vision` tool that describes and OCRs image files by calling the free Zhipu GLM vision API directly (no external CLI required).
dsh plugin add dsh-vision-free-eyesDSH 静默视觉增强:主模型照常选择,图片自动交给固定视觉模型后以隐藏上下文返回主模型。
dsh plugin add dsh-vision-fallbackBridges images into text for DeepSeek Harness: when the session's selected model cannot see images, a configured vision model describes them and the descriptions enter the durable session history as folded context rows — your message stays untouched.
dsh plugin add dsh-auto-imageGive DeepSeek Harness free eyes — natively. Paste or drop an image into the Web GUI composer: it appears as a thumbnail attachment (like a normal AI chat), and before the text-only DeepSeek model is called, the host automatically reads the image with the
dsh plugin add dsh-dseyes为 DeepSeek Harness 提供外部视觉模型能力:纯文本主模型通过 describe_image 工具调用外部视觉模型看图,拿到纯文字描述(多模态补全)。静态 Cordis 插件,随 DSH 启动自动加载。
dsh plugin add @woyeshishen/dsh-vision-pluginDeepSeek Harness 视觉桥:自动发现你已配置的多模态模型,给纯文本主模型装上 vision 工具,识别结果以纯文本返回。零配置,一条命令安装。
dsh plugin add dsh-auto-visionAdaptive image routing for DeepSeek Harness: pass images directly to native multimodal models and transcribe them only for text-only models.
dsh plugin add dsh-vision-recognizerDeepSee: DeepSeek Harness vision, multi-model routing, and Gemini visual self-checks before delivery
dsh plugin add @chang416/deepseeMindsEye: model-driven vision tools, structured evidence, and exact cache for DeepSeek Harness
dsh plugin add dsh-mindseyeVision toolkit for DeepSeek Harness -- glance, ground, detect, crop CLI tools for text-only agents to understand images
dsh plugin add dsh-plugin-vision-toolkitDSH Web 插件:调用智谱 GLM 视觉模型分析图片(analyze_image 工具 + 插件设置卡片)。
dsh plugin add @gelomen/glm-vision-pluginDeepSeek Harness plugin: Bilibili keyword video search, video metadata, subtitle transcripts, direct play URLs, and multimodal frame viewing (bilibili_search / bilibili_video / bilibili_subtitles / bilibili_playurl / bilibili_frames). Anonymous by default
dsh plugin add dsh-plugin-bilibilidsh bundle: subagent_vision — delegate image reading to a vision-capable model from a text-only session, plus paste-to-path so pasted images reach the subagent as file paths.
dsh plugin add dsh-subagent-visionIntelligentEyes universal agent vision gateway for DeepSeek Harness
dsh plugin add @intelligenteyes/dshLightweight, route-preserving vision link for DSH: a configured vision model sees while the selected text model stays in control.
dsh plugin add dsh-vision-linkVision for text-only DeepSeek Harness agents: delegate image reading to a one-shot subagent on a configurable vision route (MiniMax / Kimi / any OpenAI-compatible provider), keeping image bytes and the vision model's context out of the main session.
dsh plugin add dsh-vision-subagentWindows computer-use capability for DeepSeek Harness: screenshot → vision model → simulated mouse/keyboard input, with self-evolving knowledge base.
dsh plugin add dsh-computer-use-visionTransparent image guard + vision analysis for DeepSeek Harness: text-only models read pasted images without the 400 session deadlock.
dsh plugin add dsh-vision-guardFully-local OCR CLI for text-only LLMs: PaddleOCR-VL first-tier engine with tesseract fallback, dsh plugin included
dsh plugin add local-ocr-cliNative dsh (DeepSeek Harness) Cordis plugin adapter for vision-translation: grounds images into <vision-context> via the Python CLI (PROTOCOL v1). Spawns cli.py, never re-implements core logic.
dsh plugin add vision-translation-dshDSH 双模型路由插件:主对话留在 DeepSeek,图片任务自动分发到显式配置的视觉模型,再把分析结果交回 DeepSeek 继续。
dsh plugin add @acc1143/dsh-vision-bridge