Skip to content

image-understanding 13

  • dsh-vision-proxy v0.4.1 Verified 15

    DeepSeek brain + automatic image transcription for DeepSeek Harness: a deepseek-vision provider route that transcribes attached images to text before delegating to the text-only DeepSeek adapter. Official deepseek-v4-flash-vision-exp by default (a pure-te

    dsh plugin add dsh-vision-proxy
  • dsh-vision-web v0.1.0 Verified Web UI 8

    Eyes for text-only DeepSeek on DeepSeek Harness: image understanding via Doubao Web by default (zero cost, no API key), Antigravity IDE quota (flash/pro), Gemini API, or Cockpit proxy — model-invokable vision tool, wrapper adapters, evidence memory, and a

    dsh plugin add dsh-vision-web
  • dsh-vision-bridge v1.0.0 Verified Web UI 6

    On-demand vision for text-only DSH sessions: images become markers, and a vision_describe tool sends only image + question to an OpenAI-compatible vision model

    dsh plugin add @dsh-extension/dsh-vision-bridge
  • dsh-codex-tools v1.0.1 Verified 5

    Codex-backed web search, image generation, and image understanding tools for the DeepSeek Harness.

    dsh plugin add dsh-codex-tools
  • dsh-youreyes v0.1.0 Verified Web UI 3

    Eyes for text-only DeepSeek on DeepSeek Harness: image understanding via Antigravity IDE quota (default, flash/pro), OpenAI-compatible VLM endpoints, Gemini API, or local Ollama — model-invokable vision tool, wrapper adapters for deepseek/opencode-go, evi

    dsh plugin add dsh-youreyes
  • dsh-vision-recognizer v0.2.0 Verified Web UI 2

    Adaptive image routing for DeepSeek Harness: pass images directly to native multimodal models and transcribe them only for text-only models.

    dsh plugin add dsh-vision-recognizer
  • dsh-vision-link v1.2.2 Verified Web UI 1

    Lightweight, route-preserving vision link for DSH: a configured vision model sees while the selected text model stays in control.

    dsh plugin add dsh-vision-link
  • dsh-mmx v1.5.0 Verified Web UI 1

    MiniMax CLI bridge for DeepSeek Harness (dsh): mmx-backed web search fallback and transparent image understanding, with first-run onboarding and a settings card.

    dsh plugin add dsh-mmx
  • dsh-image-plugins v0.1.4 Verified 1

    Multimodal plugin for DeepSeek Harness: understand images and generate images through configurable OpenAI-compatible or DashScope endpoints.

    dsh plugin add dsh-image-plugins
  • dsh-free-vision v0.1.2 Verified Web UI

    Local image understanding for the DeepSeek Harness: OCR, table-layout detection, and semantic description of textless images via macOS Vision, fully on-device (images never leave your Mac). 本地化识图,图片不出本机。

    dsh plugin add @freespace8/dsh-free-vision
  • dsh-plugin-vision v0.1.0 Verified

    Lets a text-only agent call a multimodal model mid-task: the vision tool sends one image file to Qwen, Kimi, OpenAI, Claude, Gemini or a self-hosted endpoint and returns structured evidence — summary, verbatim OCR, layout, entities, and what the model cou

    dsh plugin add dsh-plugin-vision
  • dsh-vision-bridge v0.1.0 Verified Web UI

    Let text-only models see images in DeepSeek Harness: intercepts the llm/stream waterfall and transparently substitutes each image with a vision model's description.

    dsh plugin add @zzdream67/dsh-vision-bridge
  • dsh-vision-local v0.3.0 Verified

    Local-first vision for text-only DeepSeek Harness agents: route image understanding to a local OpenAI-compatible vision model, return structured JSON evidence (summary / OCR / layout / semantics / visual / uncertainty).

    dsh plugin add dsh-vision-local