Screen capture and external vision recognition: take_screenshot, list_windows, analyze_image and view_image tools with a configurable GPT vision channel (gpt-5.5 / gpt-5.6-sol / gpt-5.6-terra), API key via the credentials service, and a settings card; view_image shows the screenshot in the Web UI while the model context keeps text only.
- License
- MIT
- Added
- 2026-08-21
GitHub info
- License
- MIT
- Primary language
- TypeScript
- Last push
- Aug 20, 2026, 3:06 PM
- Maintainer
- ankye
- Added
- 2026-08-21
Install
dsh plugin --profile web add github:ankye/dsh-client-vision#path:/packages/tool-visionREADME badge
Add this Markdown to your plugin README to link back to its listing.
[](https://dshget.com/plugins/ankye/dsh-client-vision%23tool-vision)Related plugins
Vision & Multimodalmodlens
★ 3768Vision bridge for text-only models: paste an image, get structured JSON evidence (OCR, layout, semantics).
dsh-vision-router
★ 1030Free vision for text-only agents: built-in keyless vision chain plus pixel tools (Q&A, grounding, crop, pixel diff, colors, OCR, SVG trace, cutout, screenshots); paste an image to use it.
dsh-vision-toolkit
★ 842Vision for text-only models: paste an image and the model switches to a Vision Toolkit variant for image Q&A, multi-image comparison, long-screenshot OCR, screenshot-to-UI reproduction, element grounding, and pixel diff. No API key by default — images are processed by the author-hosted free service, 100 per machine per day; configurable to your own provider.