概览
dsh-tool-vision
README / ZH
插件文档
目录摘要
dsh.pub 核对固定版本的组合包契约、运行时事实与分发语义;完整 README 请查看源仓库。
在 GitHub 阅读完整 READMELIMITATIONS
已知限制
- A bridged image enters the conversation as a text hint (a transcript, not pixels) — pixel-precise in-context reasoning is not available to text-only models; the vision model's description comes back through `inspect_image`. - The bridge is a **one-way door**: an image pasted on a text-only model is rewritten into the durable log at `agent/pre-step`, so switching to a multimodal model later does not turn it back into an image block. (The other direction — multimodal to text-only — is repaired automatically by `repairLoggedImages`.) - Images are base64-transferred; mind privacy and size limits. - Independent of the dsh-llm routing/retry system; failures return clear errors to the agent.
