Hyp6666/dsh-open-eyes
Hyp6666★ 1TypeScriptLast synced: 2026-08-15
A lightweight DeepSeek Harness vision delegation tool for text-only routes, with native OpenAI Responses, Chat Completions, and Anthropic Messages adapters.
README excerpt
dsh-open-eyes Let the multimodal model you choose become DeepSeek's eyes. English · 中文 What it does The main model used by DeepSeek Harness does not always support images. When a conversation involves a screenshot, photo, chart, or interface, dsh-open-eyes can send the image to a separately configured multimodal model and return its analysis as text to the current conversation. If the current main model already supports images, the plugin stays out of the way and DSH keeps using its native image path. Images pasted, dropped, or selected in the WebUI are bridged only when the current model is explicitly known not to support images. The plugin also provides the vision analyze tool for analyzing local image paths and explicitly enabled remote image URLs. Unofficial community plugin: dsh-open-eyes is an independent community project. It is not affiliated with, endorsed by, or maintained by DeepSeek. Three APIs are supported: - OpenAI Responses - OpenAI Chat Completions - Anthropic Messages Any vision service that implements one of these APIs can be connected. You choose the endpoint, model, and credentials; the plugin is not tied to a particular provider. API keys are resolved through …
View full README on GitHub →Category
The first vision plugin for DeepSeek Harness, and the vision bridge for every text-only coding agent. Paste an image, get structured JSON evidence (OCR, layout, semantics). | 全网第一个 DeepSeek Harness 视觉插件,为 DeepSeek、GLM 等纯文本模型外挂视觉能力,粘贴图片即得结构化 JSON 证据(OCR、版面、语义)。
★ 1,700
Anionex/agent-vision-toolkit为纯文本模型"看图“设计更好的视觉工具箱和技能,支持多图理解,图片问答,前端UI还原、GUI 自动化等,并可选无缝接入多个主流agent,直接识别粘贴图片| A vision toolkit and skill designed for text-only llms — image Q&A, long-screenshot OCR, frontend UI restoration, and GUI automation, with optional seamless integration for Codex, Claude Code, Pi, Oh My Pi, and OpenCode
★ 884
Anionex/dsh-vision-toolkit让纯文本模型更好地做视觉任务的DeepSeek Harness插件:带意图的图片问答、长截图 OCR、UI 还原等|DeepSeek Harness-native integration for agent-vision-toolkit: image Q&A, long-screenshot OCR, UI restoration, grounding, pixel diff, Artifacts, and Web UI.
★ 405