为纯文本模型架起视觉桥梁:粘贴图片,输出结构化 JSON 证据(OCR、版面、语义)。
- GitHub 星标
- 3,445
- 许可证
- MIT
- 收录日期
- 2026-08-14
GitHub 信息
- GitHub 星标
- 3,445
- 许可证
- MIT
- 主要语言
- TypeScript
- 最近推送
- 2026年8月20日 20:05
- 维护者
- liustack
- 收录日期
- 2026-08-14
README 演示图
从仓库 README 提取的截图/GIF(已过滤徽章、头像等装饰图)。

The skill triggering on its own in a DeepSeek Claude Code session and reading a pasted slide
https://raw.githubusercontent.com/liustack/modlens/main/assets/demo-claude-paste-recovery.jpg

Text-only DeepSeek reading a tweet screenshot in full detail via ModLens
https://raw.githubusercontent.com/liustack/modlens/main/assets/demo-codex-app.jpg

Three images dropped together, read one by one
https://raw.githubusercontent.com/liustack/modlens/main/assets/demo-codex-batch.jpg

The 128-model scatter plot read in full: axes, log scale, and highlighted region
https://raw.githubusercontent.com/liustack/modlens/main/assets/demo-codex-chart.jpg

Pasting an image straight into DeepSeek Harness, read through the modlens vision plugin
https://raw.githubusercontent.com/liustack/modlens/main/assets/demo-dsh-paste.jpg

The ModLens vision-engine card in the dsh settings page, shown in Chinese: switch the engine, tick which local CLIs auto mode reuses
https://raw.githubusercontent.com/liustack/modlens/main/assets/demo-dsh-settings-card.jpg
安装
dsh plugin --profile web add @liustack/modlens标签
README 徽章
将这段 Markdown 添加到插件 README,链接到对应的详情页。
[](https://dshget.com/plugins/liustack/modlens)相关插件
视觉与多模态modsearch
★ 321纯文本 agent 的联网搜索桥:搜索网页与 X,返回结构化 JSON 证据(search/fetch/引用)。
agent-vision-toolkit
★ 1121为纯文本模型"看图“设计更好的视觉工具箱和技能,支持多图理解,图片问答,前端UI还原、GUI 自动化等,并可选无缝接入多个主流agent,直接识别粘贴图片| A vision toolkit and skill designed for text-only llms — image Q&A, long-screenshot OCR, frontend UI restoration, and GUI automation, with optional seamless integration for Codex, Claude Code, Pi, Oh My Pi, and OpenCode
dsh-vision-toolkit
★ 842让纯文本模型处理视觉任务:粘贴图片后自动切换到 Vision Toolkit 变体,支持图片问答、多图比较、长截图 OCR、截图还原前端 UI、元素定位与像素对比。默认无需 API Key——图片经作者自建的免费服务处理,每台机器每天 100 张;也可改为指向自己的服务商。