跳到主要内容

multimodal 75

  • dsh-image-guard v0.8.17 已验证 Web 界面

    DeepSeek Harness plugin: trims the oldest images from outgoing chat requests and, when the provider rejects the image count with HTTP 400, parses the limit from the error and retries with fewer images. · DSH 插件:裁剪即将发出请求中的历史图片,并在上游因图片数量返回 400 时解析该上限并按更少的图片

    dsh plugin add dsh-image-guard
  • dsh-agnes-multimodal v1.0.0 已验证 Web 界面

    Agnes AI 多模态融合插件:一个包同时提供「多模态配置」设置页(账号池,写入 $DSH_HOME/ag-multimodal.json)、generate_image 图片生成工具与 generate_video 视频生成工具,以及三者共用的一套会话内卡片 UI。融合自 dsh-agconfig + dsh-agimage + dsh-agvideo。

    dsh plugin add dsh-agnes-multimodal
  • dsh-image-reader v0.1.1 已验证

    让纯文本会话模型也能读取用户上传的图片——包装原生 deepseek-official 适配器:声明图片输入能力放行附件,把图片交给视觉模型转成文字,再委托给纯文本 DeepSeek。

    dsh plugin add @cxxl/dsh-image-reader
  • dsh-mingmu v0.1.5 已验证 Web 界面

    明眸 VisionBridge - 自研视觉桥:瞎子模型收图时自动调用视觉模型识别,把识别文字喂回主模型,无感、可配置、升级不丢

    dsh plugin add dsh-mingmu
  • deepseek-vision-bridge v0.1.3 已验证 Web 界面

    为 DSH 非视觉模型提供基于 DeepSeek V4 Flash Vision Exp 的会话图片分析工具

    dsh plugin add @p-dsh-market/deepseek-vision-bridge
  • dsh-media-guard v0.1.0-alpha.2 已验证 Web 界面

    Request image optimization, intelligent retention, and observability for DeepSeek Harness (DSH)

    dsh plugin add dsh-media-guard
  • useful-dsh-plugins v0.4.2 已验证

    Meta package: one-command installation of useful community plugins for DeepSeek Harness — upload button, document reader, vision (image reading via DeepSeek's built-in multimodal model), plugin manager, and desktop launcher.

    dsh plugin add useful-dsh-plugins
  • dsh-eyes v0.4.0 已验证 Web 界面

    A vision bridge profile bundle for DeepSeek Harness (dsh): gives non-vision models image-reading capability by delegating transcription to a vision-capable model, with GUI paste admission and a web settings page for vision model selection.

    dsh plugin add dsh-eyes
  • dsh-bundle-vision v0.1.0 已验证

    Vision bundle + plugin for DeepSeek Harness: the describe_image tool reads local images and asks any configured multimodal route, with zero core changes

    dsh plugin add dsh-bundle-vision
  • dsh-provider v0.1.6 已验证

    TokenLab provider bundle for DeepSeek Harness with native Responses, Messages and Chat routing plus multimodal and async tools.

    dsh plugin add @tokenlabai/dsh-provider
  • dsh-vision-bridge v0.1.12 已验证 Web 界面

    DSH 视觉桥插件:自动区分多模态/文本模型。多模态模型直接看图;文本模型通过可配置的多模态端点(baseUrl + apiKey + model)代看,支持粘贴图片、read_image 工具、请求时图片转证据。

    dsh plugin add @omdp/dsh-vision-bridge
  • intelligenteyes v2.0.0 已验证 Web 界面

    Native-vision capability manager and lightweight external vision fallback for DeepSeek Harness

    dsh plugin add @starnight11123/intelligenteyes
  • dsh-tool-vision v0.1.2 已验证 Web 界面

    Vision tools (analyze_image, locate_element) for DeepSeek Harness: local ollama, DSH subagent (qwen-vl etc), or OpenAI-compatible HTTP endpoints

    dsh plugin add @bujue3184/dsh-tool-vision
  • dsh-vision-handoff v0.1.0 已验证 Web 界面

    Describe images with a vision model so text-only DeepSeek Harness models can read them

    dsh plugin add dsh-vision-handoff
  • dsh-vision-bridge v0.1.0 已验证 Web 界面

    Let text-only models see images in DeepSeek Harness: intercepts the llm/stream waterfall and transparently substitutes each image with a vision model's description.

    dsh plugin add @zzdream67/dsh-vision-bridge
  • dsh-vision-pro-bridge v1.0.0 已验证

    Give text-only DeepSeek-V4-Pro real vision with zero new dependencies and DeepSeek-only routing: images are described by deepseek-v4-flash-vision-exp (your existing DEEPSEEK_API_KEY), then the text is handed to V4-Pro.

    dsh plugin add dsh-vision-pro-bridge
  • dsh-vision-helper v0.4.2 已验证 Web 界面

    Persistent vision plugin for DeepSeek Harness: registers the vision_analyze tool backed by a configurable multimodal model, with a settings-page UI. Zero dependencies.

    dsh plugin add dsh-vision-helper
  • dsh-vision-worker v0.1.1 已验证

    DeepSeek Harness plugin: a vision worker over Cloudflare Workers AI (@cf/moonshotai/kimi-k2.6) that routes image requests from text-only callers, returns a versioned righthand.vision.v1 envelope, and supports follow-up questions.

    dsh plugin add @try-works/dsh-vision-worker
  • dsh-vision-guard v0.1.3 已验证

    Transparent image guard + vision analysis for DeepSeek Harness: text-only models read pasted images without the 400 session deadlock.

    dsh plugin add dsh-vision-guard
  • dsh-vision-adapter v0.6.4 已验证 Web 界面

    给 DeepSeek Harness 加视觉能力:可视化设置页选视觉厂商(Kimi/智谱/通义/OpenAI/Gemini/豆包/MiniMax/阶跃星辰)并粘贴 API Key,聊天里即可拖图识别——图片走视觉模型、文字走 DeepSeek 推理。

    dsh plugin add dsh-vision-adapter
  • dsh-image-gen v0.2.0 已验证 Web 界面

    DeepSeek Harness 生图插件:侧边栏全局画廊(Cherry Studio 式绘画页)、对话生图/编辑、可自定义 Provider 端点与代理。

    dsh plugin add @copylee/dsh-image-gen
  • dsh-plugin-vision v0.1.0 已验证

    Lets a text-only agent call a multimodal model mid-task: the vision tool sends one image file to Qwen, Kimi, OpenAI, Claude, Gemini or a self-hosted endpoint and returns structured evidence — summary, verbatim OCR, layout, entities, and what the model cou

    dsh plugin add dsh-plugin-vision
  • dsh-tool-vision v0.1.3 已验证

    DSH plugin: read images through a user-configured vision-capable model when the main model route cannot accept image input

    dsh plugin add @pzqian123/dsh-tool-vision
  • dsh-plugin-mm-vision v0.1.1 已验证

    mm-vision (通感编码器) for DeepSeek Harness — give any text-only LLM the ability to see images via structured spatial text encoding. Registers the mm_vision tool.

    dsh plugin add dsh-plugin-mm-vision
  • dsh-plugin-image-tools v0.6.8 已验证 Web 界面

    DSH 图片插件,三个工具覆盖三种场景:ask_user_choice 图片/图文混合选择卡(Web GUI 渲染,可放大查看)+ show_images 回复内嵌图片(图文混排)+ save_received_images 盲模型收图存为工作区文件;聊天栏所有图片点击放大,支持滚轮缩放与拖拽平移。来源支持本地路径 / http(s) URL / base64 data URI。零 token 本地渲染,纯插件实现不改核心包。

    dsh plugin add dsh-plugin-image-tools