Skip to content

ocr 43

  • modlens v3.25.0 Verified Web UI 3695

    Plug-in vision for text-only LLMs, powered by the free Antigravity CLI

    dsh plugin add @liustack/modlens
  • dsh-vision-router v2.0.1 Verified Web UI 997

    Eyes for text-only DeepSeek Harness agents: built-in free vision chain (no key) + pixel-level vision tools (Q&A, grounding, crop, pixel diff, colors, OCR, SVG trace, cutout, screenshots). One-command install, no Python, image turns work like ordinary tool

    dsh plugin add dsh-vision-router
  • dsh-vision-toolkit v0.1.39 Verified Web UI 834

    DeepSeek Harness-native integration for agent-vision-toolkit: image Q&A, OCR, grounding, UI restoration, pixel diff, Artifacts, and Web UI.

    dsh plugin add @anionex/dsh-vision-toolkit
  • dsh-invoice-downloader v0.1.1 Verified Web UI 141

    Local IMAP invoice download, OCR, archive, and Excel summary bundle for DeepSeek Harness

    dsh plugin add @ethanyoq/dsh-invoice-downloader
  • picturereader v3.2.0 Verified Web UI 34

    Unified image understanding plugin for DeepSeek Harness (DSH). Visual twin adapter for native thumbnails + auto-analysis on any text-only model (incl. pi-ai providers); privacy/smart/strict routing; local tools (scan/OCR×3 engines/crop/palette/compare/bat

    dsh plugin add picturereader
  • dsh-vision-proxy v0.4.1 Verified 15

    DeepSeek brain + automatic image transcription for DeepSeek Harness: a deepseek-vision provider route that transcribes attached images to text before delegating to the text-only DeepSeek adapter. Official deepseek-v4-flash-vision-exp by default (a pure-te

    dsh plugin add dsh-vision-proxy
  • dsh-doc v0.1.1 Verified 13

    Local PDF, Office, image, and OCR document intelligence for DeepSeek Harness.

    dsh plugin add dsh-doc
  • dsh-chat-imagine v0.4.1 Verified 11

    Generate and display images inline in the DSH chat via API channels or local CLIs (mmx / codex / agy), and read images into structured JSON evidence (OCR / layout / semantics) on any model — with backend probing and a bundled recovery skill.

    dsh plugin add dsh-chat-imagine
  • dsh-vision-web v0.1.0 Verified Web UI 8

    Eyes for text-only DeepSeek on DeepSeek Harness: image understanding via Doubao Web by default (zero cost, no API key), Antigravity IDE quota (flash/pro), Gemini API, or Cockpit proxy — model-invokable vision tool, wrapper adapters, evidence memory, and a

    dsh plugin add dsh-vision-web
  • deepseekeyes v0.8.1 Verified Web UI 7

    Auditable vision and cross-platform Computer Use runtime for DeepSeek Harness with source-preserving evidence.

    dsh plugin add @dttxorg/deepseekeyes
  • dsh-windows-ocr v0.2.6 Verified 7

    dsh plugin: recognize attached images with the built-in Windows OCR engine (Windows.Media.Ocr) and send only the recognized text to the model. Text models never receive image bytes; vision passthrough is opt-in.

    dsh plugin add @maxwell-feng/dsh-windows-ocr
  • dsh-ocr-local v0.3.3 Verified Web UI 6

    Local OCR for DeepSeek Harness: paste/attach an image, get its text via PP-OCRv5 + ONNX Runtime, fully offline. TUI (cc-tui) and Web. / DeepSeek Harness 本地 OCR 插件:图片转文字,PP-OCRv5 + ONNX Runtime,完全离线,支持 TUI 与 Web。

    dsh plugin add dsh-ocr-local
  • dsh-vision-bridge v1.0.0 Verified Web UI 6

    On-demand vision for text-only DSH sessions: images become markers, and a vision_describe tool sends only image + question to an OpenAI-compatible vision model

    dsh plugin add @dsh-extension/dsh-vision-bridge
  • dsh-plugin-deepseek-vision v0.4.1 Verified Web UI 5

    DeepSeek Harness 原生视觉 Bundle:粘贴或拖入图片,通过托管的 deepseek-vision-mcp 调用 OpenAI 兼容视觉模型。

    dsh plugin add dsh-plugin-deepseek-vision
  • dsh-plugin-deepeye v0.2.0 Verified 4

    DeepSeek Harness native plugin: vision capabilities for text-only LLMs (describe, OCR, VQA, layout analysis, clipboard)

    dsh plugin add dsh-plugin-deepeye
  • dsh-mineru v0.1.9 Verified Web UI 4

    dsh-mineru: MinerU 文档解析插件 for DeepSeek Harness — 多模态全格式 (PDF/Word/PPT/Excel/HTML/图片) → 结构化 Markdown. 填 Token 走精准解析 API, 留空走 Agent 轻量解析 API.

    dsh plugin add dsh-mineru
  • dsh-tesseract-ocr v0.2.6 Verified 3

    dsh plugin: recognize attached images locally with Tesseract OCR and send only the recognized text to the model. Text models never receive image bytes; vision passthrough is opt-in.

    dsh plugin add @maxwell-feng/dsh-tesseract-ocr
  • dsh-open-file v0.1.1-rc.2 Verified Web UI 3

    Workspace-bound arbitrary file upload, reading, OCR, and rendering for DeepSeek Harness.

    dsh plugin add dsh-open-file
  • dsh-youreyes v0.1.0 Verified Web UI 3

    Eyes for text-only DeepSeek on DeepSeek Harness: image understanding via Antigravity IDE quota (default, flash/pro), OpenAI-compatible VLM endpoints, Gemini API, or local Ollama — model-invokable vision tool, wrapper adapters for deepseek/opencode-go, evi

    dsh plugin add dsh-youreyes
  • dsh-ui-spec v0.1.1 Verified 3

    DeepSeek Harness plugin that turns UI screenshots into implementation-grade web specs using OCR, deterministic geometry, scene graphs, assets, and render comparison.

    dsh plugin add dsh-ui-spec
  • dsh-vision-free-eyes v0.1.3 Verified 2

    DeepSeek Harness plugin: a model-facing `vision` tool that describes and OCRs image files by calling the free Zhipu GLM vision API directly (no external CLI required).

    dsh plugin add dsh-vision-free-eyes
  • dsh-plugin-qwen-image v0.3.1 Verified Web UI 2

    DeepSeek Harness out-of-tree plugin: give a text-only coding model eyes by routing images to a Qwen-VL (DashScope) route through ctx.llm and returning text.

    dsh plugin add dsh-plugin-qwen-image
  • dsh-auto-image v0.1.1 Verified 2

    Bridges images into text for DeepSeek Harness: when the session's selected model cannot see images, a configured vision model describes them and the descriptions enter the durable session history as folded context rows — your message stays untouched.

    dsh plugin add dsh-auto-image
  • dsh-vision-recognizer v0.2.0 Verified Web UI 2

    Adaptive image routing for DeepSeek Harness: pass images directly to native multimodal models and transcribe them only for text-only models.

    dsh plugin add dsh-vision-recognizer
  • dsh-dseyes v0.2.0 Verified 2

    Give DeepSeek Harness free eyes — natively. Paste or drop an image into the Web GUI composer: it appears as a thumbnail attachment (like a normal AI chat), and before the text-only DeepSeek model is called, the host automatically reads the image with the

    dsh plugin add dsh-dseyes
  • deepsee v4.0.2 Verified Web UI 2

    DeepSee: DeepSeek Harness vision, multi-model routing, and Gemini visual self-checks before delivery

    dsh plugin add @chang416/deepsee
  • local-ocr-cli v0.1.10 Verified Web UI 1

    Fully-local OCR CLI for text-only LLMs: PaddleOCR-VL first-tier engine with tesseract fallback, dsh plugin included

    dsh plugin add local-ocr-cli
  • dsh-vision-guard v0.1.3 Verified 1

    Transparent image guard + vision analysis for DeepSeek Harness: text-only models read pasted images without the 400 session deadlock.

    dsh plugin add dsh-vision-guard
  • dsh-plugin-vision-toolkit v0.1.1 Verified 1

    Vision toolkit for DeepSeek Harness -- glance, ground, detect, crop CLI tools for text-only agents to understand images

    dsh plugin add dsh-plugin-vision-toolkit
  • dsh-autovision v0.1.0 Verified Web UI 1

    dsh auto-vision bridge: paste an image into a text-only model's composer and a configured multimodal model transcribes it as text, automatically.

    dsh plugin add @iroam2375/dsh-autovision
  • dsh-vision-link v1.2.2 Verified Web UI 1

    Lightweight, route-preserving vision link for DSH: a configured vision model sees while the selected text model stays in control.

    dsh plugin add dsh-vision-link
  • dsh-plugin-vision v0.1.0 Verified

    Lets a text-only agent call a multimodal model mid-task: the vision tool sends one image file to Qwen, Kimi, OpenAI, Claude, Gemini or a self-hosted endpoint and returns structured evidence — summary, verbatim OCR, layout, entities, and what the model cou

    dsh plugin add dsh-plugin-vision
  • free-vision-skill v0.3.1 Verified Web UI

    DSH-Plugin for DeepSeek-Harness: fully-local image understanding & OCR powered by macOS Vision Framework

    dsh plugin add @niyongsheng/free-vision-skill
  • dsh-img v0.4.0 Verified Web UI

    Give text-only models eyes: an analyze_image tool for DeepSeek Harness, backed by free Chinese vision APIs (GLM-4V-Flash / Qwen-VL) or any OpenAI-compatible vision endpoint. 给纯文本模型装上眼睛。

    dsh plugin add dsh-img
  • dsh-maclens v0.1.2 Verified

    Bridge Apple's on-device Vision framework (macOS) into DeepSeek Harness: OCR, image classification, face detection, and document layout as local dsh tools. No network, no API key, no daemon.

    dsh plugin add dsh-maclens
  • dsh-ocr-bridge v0.1.1 Verified

    Paste images into DeepSeek Harness chat and have them read by a free local backend (macOS Vision / Tesseract) before the text-only DeepSeek model answers

    dsh plugin add dsh-ocr-bridge
  • dsh-pdf-reader v0.2.0 Verified

    DeepSeek Harness plugin: content-aware PDF reading tools for vision models (pdf_scan / pdf_read_page / pdf_render_region). Detects figures (vector+raster), tables, math, and two-column layout per page, then routes text pages to extraction and figure/table

    dsh plugin add dsh-pdf-reader
  • dsh-file-attach v0.1.0 Verified Web UI

    Drag-and-drop PDF, Office, images, and text/code files into DSH conversations. The host extracts (and OCRs) them into the prompt; attach_* tools cover notebook cells, PDF-page OCR, image describe, and save.

    dsh plugin add @lucasxingg/dsh-file-attach
  • dsh-bundle-vision v0.1.0 Verified

    Vision bundle + plugin for DeepSeek Harness: the describe_image tool reads local images and asks any configured multimodal route, with zero core changes

    dsh plugin add dsh-bundle-vision
  • dsh-siliconflow-vision v1.1.1 Verified Web UI

    DSH 插件:通过硅基流动(SiliconFlow)视觉大模型识别/分析图片,支持本地文件路径、http(s) 图片 URL 与 base64 data URL。含持久化的粘贴识别面板(web 端)。

    dsh plugin add dsh-siliconflow-vision
  • dsh-vision-helper v0.4.2 Verified Web UI

    Persistent vision plugin for DeepSeek Harness: registers the vision_analyze tool backed by a configurable multimodal model, with a settings-page UI. Zero dependencies.

    dsh plugin add dsh-vision-helper
  • dsh-vision-local v0.3.0 Verified

    Local-first vision for text-only DeepSeek Harness agents: route image understanding to a local OpenAI-compatible vision model, return structured JSON evidence (summary / OCR / layout / semantics / visual / uncertainty).

    dsh plugin add dsh-vision-local
  • dsh-free-vision v0.1.2 Verified Web UI

    Local image understanding for the DeepSeek Harness: OCR, table-layout detection, and semantic description of textless images via macOS Vision, fully on-device (images never leave your Mac). 本地化识图,图片不出本机。

    dsh plugin add @freespace8/dsh-free-vision