Multimodal bridge for any LLM
Spirit4471/multimodal-bridge · model-adapters
Multi-modal bridge for any LLM: qwenvision (Qwen-VL)+ qwengerate (Qwen-Image) - MCP server for Kimi Code and Claude Code, as well as DeepSeek Harness plugin (dsh-multimedial-bridge).
★ 3GitHub stars
0Forks
2026-08-20Last updated
TypeScriptLanguage
MITLicense
Key features
- MCP Bridges — Add visual understanding and image generation to pure LLM or add the image generation that it does not have to do with a model.
- The main line is suitable for daily heavyness: K3 carries its own high-performance reasoning and vision, multimodular information in one hand, no intermediate model to reproduce losses
- Bridge will only be available when required.
- Alternative lines are suitable for a cost-sensitive scenario: DeepSeek is extremely cheap token, but it has no vision, and qwen vision is the only way it looks at the picture.
Install command
dsh plugin --profile web add github:Spirit4471/multimodal-bridge