跳到主要内容

vision 92

  • dsh-autovision v0.1.0 已验证 Web 界面 1

    dsh auto-vision bridge: paste an image into a text-only model's composer and a configured multimodal model transcribes it as text, automatically.

    dsh plugin add @iroam2375/dsh-autovision
  • dsh-gemini-multimodal v0.1.2 已验证 1

    DeepSeek Harness plugin: multimodal tools (image/audio/video/document understanding, transcription, image generation) via Gemini API or the local Antigravity CLI.

    dsh plugin add dsh-gemini-multimodal
  • dsh-glm-vision v1.0.1 已验证 Web 界面 1

    GLM 视觉模型插件:注册 glm-vision 供应商路由(glm-4.6v 系列,image+text 输入声明)并提供 glm_vision 工具,让 DeepSeek 等文本主模型直接调用智谱视觉模型看图。

    dsh plugin add dsh-glm-vision
  • dsh-image-plugins v0.1.4 已验证 1

    Multimodal plugin for DeepSeek Harness: understand images and generate images through configurable OpenAI-compatible or DashScope endpoints.

    dsh plugin add dsh-image-plugins
  • dsh-tool-visual-primitives v1.3.0 已验证 Web 界面 1

    DSH 视觉原语工具:参考 DeepSeek《Thinking with Visual Primitives》论文,将图片路由到外部视觉模型并返回带视觉基元的文本分析。纯文本循环,对话模型无需原生视觉能力即可'看见'图片。

    dsh plugin add dsh-tool-visual-primitives
  • dsh-sight v0.1.11 已验证 Web 界面 1

    DeepSeek Harness plugin: direct multimodal image transfer declarations + per-session image clearing (surface replace), with a settings page and composer controls.

    dsh plugin add @eric.wen/dsh-sight
  • dsh-local-vision v0.2.1 已验证 1

    DeepSeek Harness plugin: a model-callable local_vision tool that describes images with a local vision model through Ollama. No cloud, the image never leaves the machine. Two tiers: fast (text/colors/summary) and detailed (full description).

    dsh plugin add dsh-local-vision
  • dsh-mingmu v0.1.5 已验证 Web 界面 1

    明眸 VisionBridge - 自研视觉桥:瞎子模型收图时自动调用视觉模型识别,把识别文字喂回主模型,无感、可配置、升级不丢

    dsh plugin add dsh-mingmu
  • dsh-mmx v1.5.0 已验证 Web 界面 1

    MiniMax CLI bridge for DeepSeek Harness (dsh): mmx-backed web search fallback and transparent image understanding, with first-run onboarding and a settings card.

    dsh plugin add dsh-mmx
  • dsh-vision-bridge v0.1.7 已验证 Web 界面

    DSH 视觉桥插件:自动区分多模态/文本模型。多模态模型直接看图;文本模型通过可配置的多模态端点(baseUrl + apiKey + model)代看,支持粘贴图片、read_image 工具、请求时图片转证据。

    dsh plugin add @omdp/dsh-vision-bridge
  • dsh-tool-vision v0.1.2 已验证 Web 界面

    Vision tools (analyze_image, locate_element) for DeepSeek Harness: local ollama, DSH subagent (qwen-vl etc), or OpenAI-compatible HTTP endpoints

    dsh plugin add @bujue3184/dsh-tool-vision
  • vision-tool v0.1.0 已验证

    Conversation-style UI/UX visual review plugin for DeepSeek Harness: vision_review / vision_ask tools backed by any OpenAI-compatible multimodal endpoint (e.g. agnes-2.5-flash)

    dsh plugin add @evan-williams/vision-tool
  • dsh-free-vision v0.1.2 已验证 Web 界面

    Local image understanding for the DeepSeek Harness: OCR, table-layout detection, and semantic description of textless images via macOS Vision, fully on-device (images never leave your Mac). 本地化识图,图片不出本机。

    dsh plugin add @freespace8/dsh-free-vision
  • dsh-vision-bridge v0.4.0 已验证 Web 界面

    Universal vision bridge for DeepSeek Harness: pick how images are processed — auto-rewrite via a vision LLM, explicit tools, or hybrid. Multi-channel endpoint (DSH catalog / OpenAI-compatible / Ollama / custom), LRU description cache, settings card with c

    dsh plugin add @goodandready/dsh-vision-bridge
  • dsh-vision-bridge v0.1.1 已验证

    DSH 视觉桥接插件:让无视觉能力的主模型看图(会话收图 + 自动转文字 + view_image 工具)

    dsh plugin add @liu__min/dsh-vision-bridge
  • free-vision-skill v0.3.1 已验证 Web 界面

    DSH-Plugin for DeepSeek-Harness: fully-local image understanding & OCR powered by macOS Vision Framework

    dsh plugin add @niyongsheng/free-vision-skill
  • deepseek-vision-bridge v0.1.3 已验证 Web 界面

    为 DSH 非视觉模型提供基于 DeepSeek V4 Flash Vision Exp 的会话图片分析工具

    dsh plugin add @p-dsh-market/deepseek-vision-bridge
  • dsh-tool-vision v0.1.3 已验证

    DSH plugin: read images through a user-configured vision-capable model when the main model route cannot accept image input

    dsh plugin add @pzqian123/dsh-tool-vision
  • dsh-vision-worker v0.1.1 已验证

    DeepSeek Harness plugin: a vision worker over Cloudflare Workers AI (@cf/moonshotai/kimi-k2.6) that routes image requests from text-only callers, returns a versioned righthand.vision.v1 envelope, and supports follow-up questions.

    dsh plugin add @try-works/dsh-vision-worker
  • dsh-vision-bridge v0.1.0 已验证 Web 界面

    Let text-only models see images in DeepSeek Harness: intercepts the llm/stream waterfall and transparently substitutes each image with a vision model's description.

    dsh plugin add @zzdream67/dsh-vision-bridge
  • deepseek-vl-support v0.2.9 已验证

    Give DeepSeek (text-only) models vision in Claude Code, Codex, and Agent Plugins clients: describe images via any OpenAI-compatible vision endpoint.

    dsh plugin add deepseek-vl-support
  • dsh-bundle-vision v0.1.0 已验证

    Vision bundle + plugin for DeepSeek Harness: the describe_image tool reads local images and asks any configured multimodal route, with zero core changes

    dsh plugin add dsh-bundle-vision
  • dsh-desktop-automation v0.1.1 已验证 Web 界面

    macOS desktop control for DeepSeek Harness: agent operates non-browser apps (CapCut/PS/WPS/native clients) like a human — see screen, move mouse, type text. 14 tools + vision closed-loop (see/locate/click/verify).

    dsh plugin add dsh-desktop-automation
  • dsh-eyes v0.4.0 已验证 Web 界面

    A vision bridge profile bundle for DeepSeek Harness (dsh): gives non-vision models image-reading capability by delegating transcription to a vision-capable model, with GUI paste admission and a web settings page for vision model selection.

    dsh plugin add dsh-eyes
  • dsh-img v0.4.0 已验证 Web 界面

    Give text-only models eyes: an analyze_image tool for DeepSeek Harness, backed by free Chinese vision APIs (GLM-4V-Flash / Qwen-VL) or any OpenAI-compatible vision endpoint. 给纯文本模型装上眼睛。

    dsh plugin add dsh-img
  • dsh-llm-capabilities v0.1.1 已验证 Web 界面

    DSH plugin: auto-detect and configure model capabilities (reasoningEfforts + input modalities) for llm-pi-ai. Successor to dsh-reasoning-efforts.

    dsh plugin add dsh-llm-capabilities
  • dsh-maclens v0.1.2 已验证

    Bridge Apple's on-device Vision framework (macOS) into DeepSeek Harness: OCR, image classification, face detection, and document layout as local dsh tools. No network, no API key, no daemon.

    dsh plugin add dsh-maclens
  • dsh-ocr-bridge v0.1.1 已验证

    Paste images into DeepSeek Harness chat and have them read by a free local backend (macOS Vision / Tesseract) before the text-only DeepSeek model answers

    dsh plugin add dsh-ocr-bridge
  • dsh-pdf-reader v0.2.0 已验证

    DeepSeek Harness plugin: content-aware PDF reading tools for vision models (pdf_scan / pdf_read_page / pdf_render_region). Detects figures (vector+raster), tables, math, and two-column layout per page, then routes text pages to extraction and figure/table

    dsh plugin add dsh-pdf-reader
  • dsh-plugin-mm-vision v0.1.1 已验证

    mm-vision (通感编码器) for DeepSeek Harness — give any text-only LLM the ability to see images via structured spatial text encoding. Registers the mm_vision tool.

    dsh plugin add dsh-plugin-mm-vision
  • dsh-plugin-vision v0.1.0 已验证

    Lets a text-only agent call a multimodal model mid-task: the vision tool sends one image file to Qwen, Kimi, OpenAI, Claude, Gemini or a self-hosted endpoint and returns structured evidence — summary, verbatim OCR, layout, entities, and what the model cou

    dsh plugin add dsh-plugin-vision
  • dsh-siliconflow-vision v1.1.1 已验证 Web 界面

    DSH 插件:通过硅基流动(SiliconFlow)视觉大模型识别/分析图片,支持本地文件路径、http(s) 图片 URL 与 base64 data URL。含持久化的粘贴识别面板(web 端)。

    dsh plugin add dsh-siliconflow-vision
  • dsh-tool-accurate-vision v0.1.0 已验证

    Model-facing accurate_vision tool: precise image spatial reasoning via a vision model

    dsh plugin add dsh-tool-accurate-vision
  • dsh-tool-describe-image v0.5.1 已验证 Web 界面

    DSH plugin: image understanding via any OpenAI-compatible vision API, paste-to-describe, and an animated whale-buddy desktop pet with status bubbles and a floating settings panel

    dsh plugin add dsh-tool-describe-image
  • dsh-vision-adapter v0.6.4 已验证 Web 界面

    给 DeepSeek Harness 加视觉能力:可视化设置页选视觉厂商(Kimi/智谱/通义/OpenAI/Gemini/豆包/MiniMax/阶跃星辰)并粘贴 API Key,聊天里即可拖图识别——图片走视觉模型、文字走 DeepSeek 推理。

    dsh plugin add dsh-vision-adapter
  • dsh-vision-helper v0.4.2 已验证 Web 界面

    Persistent vision plugin for DeepSeek Harness: registers the vision_analyze tool backed by a configurable multimodal model, with a settings-page UI. Zero dependencies.

    dsh plugin add dsh-vision-helper
  • dsh-vision-local v0.3.0 已验证

    Local-first vision for text-only DeepSeek Harness agents: route image understanding to a local OpenAI-compatible vision model, return structured JSON evidence (summary / OCR / layout / semantics / visual / uncertainty).

    dsh plugin add dsh-vision-local
  • dsh-vision-no-vision v0.3.1 已验证

    Gives a text-only LLM vision capability

    dsh plugin add dsh-vision-no-vision
  • dsh-vision-proxy-route v0.1.3 已验证

    DeepSeek Harness plugin: a configurable provider route that transcribes pasted images via free Zhipu GLM vision models before delegating to a text-only adapter.

    dsh plugin add dsh-vision-proxy-route
  • dsh-vision-router-inline v0.1.1 已验证 Web 界面

    Display companion for dsh-vision-router: keep Auto Vision routing, and put a square picture button on each original model row. Does nothing unless dsh-vision-router is installed.

    dsh plugin add dsh-vision-router-inline
  • dsh-vision-tools v0.1.2 已验证 Web 界面

    DeepSeek Harness 视觉能力全家桶:vision_understand 工具(OpenAI 兼容视觉 API,默认免费智谱 GLM-4.6V-Flash,限流自动降级 GLM-4V)+ 粘贴/拖拽/按钮三入口识图

    dsh plugin add dsh-vision-tools
  • useful-dsh-plugins v0.4.0 已验证

    Meta package: one-command installation of useful community plugins for DeepSeek Harness — upload button, document reader, vision (image reading via DeepSeek's built-in multimodal model), plugin manager, and desktop launcher.

    dsh plugin add useful-dsh-plugins