Image-to-text input for the Web UI: paste or drag an image and it is transcribed into structured text and sent, giving text-only LLMs image-input takeover (OpenAI-compatible vision API).
- License
- MIT
- Added
- 2026-08-15
GitHub info
- License
- MIT
- Primary language
- JavaScript
- Last push
- Aug 15, 2026, 3:58 AM
- Maintainer
- Elohia
- Added
- 2026-08-15
Install
dsh plugin --profile web add dsh-plugin-image-inputREADME badge
Add this Markdown to your plugin README to link back to its listing.
[](https://dshget.com/plugins/Elohia/dsh-plugin-image-input)Related plugins
Vision & Multimodaldsh-plugin-mm-vision
★ 2Synesthesia Encoder for DSH: a vision model translates images into compact structured spatial text (canvas/elements/percentage coordinates), giving text-only LLMs pixel-level image understanding via the `mm_vision` tool.
modlens
★ 3768Vision bridge for text-only models: paste an image, get structured JSON evidence (OCR, layout, semantics).
dsh-vision-router
★ 1030Free vision for text-only agents: built-in keyless vision chain plus pixel tools (Q&A, grounding, crop, pixel diff, colors, OCR, SVG trace, cutout, screenshots); paste an image to use it.