DSH Plugins
返回列表
🤖

dsh-image-bridge

模型推理 更新于 2026.08.18

在终端中运行以下命令:

dsh plugin install haitang1/dsh-image-bridge

将以下提示词粘贴到 DeepSeek Harness 对话框中:

在 DSH 源码 checkout 根目录执行 dsh plugin install haitang1/dsh-image-bridge 即可从 https://github.com/haitang1/dsh-image-bridge 安装本插件,安装后会自动加入当前 profile 的 bundles 列表并生效。

插件介绍

Pasting a screenshot into DeepSeek Harness to ask a text-only model to read it usually ends with a flat refusal: the api-proxy rejects any image-bearing message when the model modality list is text-only. dsh-image-bridge closes exactly that gap, letting pure-text models inside the DeepSeek Harness still accept images and route them to your existing vision tools.

Under the hood the plugin patches three points. It wraps the model-info pre-check so image messages are no longer blocked at the proxy layer. It reads the image bytes via the attachments API and writes them into the session workspace .attachments/ directory. It then uses the DSH surface-replace mechanism to swap the image block in the model-visible transcript for a lightweight placeholder like [Image 1]:"" while the human-facing chat still renders the original thumbnail (clickable, zoomable) instead of a bare file path. The model can hand that placeholder to an existing vision or MCP tool—vision_glance, mcp__mcp-vision__analyze_image, mcp__mcp-vision__ocr_extract, etc.—for actual recognition. Models that natively support images are left completely untouched.

If your daily workflow is code review, doc drafting, or research with a DeepSeek-class text model and you occasionally need it to look at a screenshot, a diagram, or a PDF page, this plugin is the missing bridge. Two caveats: do not mix slash-commands and images in the same submit, and consider clearing or compressing chat history before switching back to a text-only model so leftover vision-model image blocks do not trip the adapter.

使用场景

  • 用 DeepSeek 文本模型分析截图中的一段报错代码
  • 在对话中粘贴架构图让模型理解布局并给出优化建议
  • 让纯文本模型读取扫描件 PDF 页面中的文字并总结要点

适合人员

  • 日常使用 DeepSeek 等纯文本模型进行代码开发与文档撰写的工程师
  • 已接入 vision 或 MCP 识别工具、但底层模型不支持图片输入的 DSH 用户
  • 需要在同一对话中兼顾图片可视化展示与模型侧识别的产品或测试人员