Dsh Tool Vision
Scorp1o117/dsh-tool-vision · model-adapters
Part of the DeepSeek Harness Enhancement Suite — Vision · Soul/Persona · Long-term Memory · Plugin Marketplace.
★ 8GitHub stars
1Forks
2026-09-07Last updated
JavaScriptLanguage
NOASSERTIONLicense
Key features
- GitHubScorp1o117/dsh-tool-vision · npm: dsh-tool-vision
- Thumbnail + lightboxclick to zoom full-screen
- Alibaba DashScope (Qwen-VL)— qwen-vl-plus, qwen-vl-max
- Content-hash cache keyed by endpoint+model+image+question (no stale answers across model switches, failures are never cached).
- Zero dependencies beyond the dsh SDK — works with any compatible endpoint: OpenAI GPT-4o, Qwen-VL (DashScope), GLM-4V (Zhipu), Moonshot, Gemini compatible endpoints, local Ollama, etc.
Requirements
- A bridged image enters the conversation as a text hint (a transcript, not pixels) — pixel-precise in-context reasoning is not available to text-only models
- the vision model's description comes back through inspect_image.
Install command
dsh plugin --profile web add github:Scorp1o117/dsh-tool-vision