概览
dsh-tool-bandit-search
README / ZH
插件文档
目录摘要
dsh.pub 核对固定版本的组合包契约、运行时事实与分发语义;完整 README 请查看源仓库。
在 GitHub 阅读完整 READMELIMITATIONS
已知限制
- **Bandit state is in-memory** and resets on every restart. Persisting it via `ctx.storage` (which dsh already exposes) is a natural next step. - **Reward is a heuristic**, not a measure of actual answer quality — it captures result count and latency, not whether the results were *relevant* or *correct*. A stronger version might score reward against whether the model's final answer actually used the returned sources. - **The model can still issue multiple `search` calls per turn** even when `thorough` is already broadening internally — the plugin optimizes strategy *per call*, not the model's own multi-call behavior. - Built and tested against dsh's developer preview; the plugin/tool APIs may change before a stable release.
