Tools
A DSH Plugin Bolts Vision Onto Text-Only Models: Screenshot OCR, Grounding and UI Restoration
`Anionex/dsh-vision-toolkit` (297 stars, created 2026-08-13, MIT) is a DeepSeek Harness-native integration that gives text-only models intent-aware image Q&A, long-screenshot OCR, UI restoration, visual grounding, pixel diff and Artifacts rendering. It is a direct answer to Harness's biggest practical gap at launch — DeepSeek's own models are text-first, so GUI-automation and screenshot-testing workflows were dead on arrival. The plugin exists two days after the harness did, which is itself the argument for the plugin architecture.
Source
↳ Follow the thread