Pixel-to-text image reading for text-only models
jing-hy/picturereader · skills
DSH plugin: pixel-to-text image reading for text-only models.
★ 37GitHub stars
4Forks
2026-08-24Last updated
JavaScriptLanguage
MITLicense
Key features
- Proxy in-situ packagingFor the provider of the selected model, use Proxy to package its adapter into "twins" and replace them in-situ (registerTwinAdapters, uninstall and restore by ctx.effect), without repeated registration.
- Every time I look at the picture, the problem faced by the model is actually the same: "What cost is this picture worth, and which route should be used to understand it?"
- Privacy mode (Privacy) - zero outbound call hard access control
- Intelligent/strict mode only uses external semantic understanding when needed and configured.
- Stream interceptioncapture the image block in the request → export to /.dsh/picturereader-vision/images/ → replace with a text path + local tool chain guidance → forward to the original adapter.
Requirements
- DSH's tools are presented in three modes: native (by default, the model can call all tools directly), code (only the model is allowed to call run_code directly, and other tools are folded into the generated SDK of run_code for calls), and both (both can be used).
- All tools of picturereader have nothing to do with mode: all can be called directly under native / both
Install command
dsh plugin --profile web add github:jing-hy/picturereader