全部插件

DSH / BUNDLE / CLIENT-UI

dsh-tool-vision

v0.9.6Scorp1o117 / dsh-tool-vision88b02624aa

可安装组合包UI 与客户端插件社区 · Topic 自动分析Web UI

概览

dsh-tool-vision

DeepSeek Harness 外置视觉模型插件:inspect_image 把本地图片或 http(s) 图片 URL 发给任意 OpenAI 兼容端点,视觉模型看图的文字回答直接带回对话;附 Web UI 设置栏。

README / ZH

插件文档

目录摘要

DeepSeek Harness 外置视觉模型插件:inspect_image 把本地图片或 http(s) 图片 URL 发给任意 OpenAI 兼容端点,视觉模型看图的文字回答直接带回对话;附 Web UI 设置栏。

dsh.pub 核对固定版本的组合包契约、运行时事实与分发语义;完整 README 请查看源仓库。

在 GitHub 阅读完整 README

LIMITATIONS

已知限制

- A bridged image enters the conversation as a text hint (a transcript, not pixels) — pixel-precise in-context reasoning is not available to text-only models; the vision model's description comes back through `inspect_image`. - The bridge is a **one-way door**: an image pasted on a text-only model is rewritten into the durable log at `agent/pre-step`, so switching to a multimodal model later does not turn it back into an image block. (The other direction — multimodal to text-only — is repaired automatically by `repairLoggedImages`.) - Images are base64-transferred; mind privacy and size limits. - Independent of the dsh-llm routing/retry system; failures return clear errors to the agent.