Overview
dsh-tool-vision
README / EN
Package documentation
Registry summary
dsh.pub verifies the pinned bundle contract, runtime facts, and distribution semantics. The complete README remains in the source repository.
Read the full README on GitHubLIMITATIONS
Known limitations
- A bridged image enters the conversation as a text hint (a transcript, not pixels) — pixel-precise in-context reasoning is not available to text-only models; the vision model's description comes back through `inspect_image`. - The bridge is a **one-way door**: an image pasted on a text-only model is rewritten into the durable log at `agent/pre-step`, so switching to a multimodal model later does not turn it back into an image block. (The other direction — multimodal to text-only — is repaired automatically by `repairLoggedImages`.) - Images are base64-transferred; mind privacy and size limits. - Independent of the dsh-llm routing/retry system; failures return clear errors to the agent.
