Synesthesia Encoder for DSH: a vision model translates images into compact structured spatial text (canvas/elements/percentage coordinates), giving text-only LLMs pixel-level image understanding via the `mm_vision` tool.
- GitHub stars
- 1
- License
- MIT
- Added
- 2026-08-15
GitHub info
- GitHub stars
- 1
- License
- MIT
- Primary language
- JavaScript
- Last push
- Aug 14, 2026, 4:25 AM
- Maintainer
- Elohia
- Added
- 2026-08-15
Install
dsh plugin --profile web add github:Elohia/dsh-plugin-mm-visionREADME badge
Add this Markdown to your plugin README to link back to its listing.
[](https://dshget.com/plugins/Elohia/dsh-plugin-mm-vision)Related plugins
Vision & Multimodaldsh-plugin-image-input
★ 1Image-to-text input for the Web UI: paste or drag an image and it is transcribed into structured text and sent, giving text-only LLMs image-input takeover (OpenAI-compatible vision API).
modlens
★ 3768Vision bridge for text-only models: paste an image, get structured JSON evidence (OCR, layout, semantics).
dsh-vision-router
★ 1030Free vision for text-only agents: built-in keyless vision chain plus pixel tools (Q&A, grounding, crop, pixel diff, colors, OCR, SVG trace, cutout, screenshots); paste an image to use it.