kaixinbaba/dsh-vision-recognizer

kaixinbaba★ 0JavaScript最后同步: 2026-08-15

在 GitHub 打开

DeepSeek Harness 识图插件:保持 DeepSeek 对话,15+ 供应商视觉模型把图片转译为文字,可在 设置→插件 配置

README 摘要

dsh-vision-recognizer English 简体中文 Keep DeepSeek as the conversation brain, attach images anyway, and switch the image-recognition provider any time from Settings → Plugins. A vision plugin for DeepSeek Harness. It registers a new provider route (default vision-recognizer , shown as DeepSeek + 识图 in the model picker) that wraps the real DeepSeek adapter: it declares image input (so the attachment preflight and the read image gate admit images) and, in the request stream, transcribes every attached image to text through the vision model you select , then delegates the text-only conversation to DeepSeek. DeepSeek still answers; recognition is an add-on. Features - One-click install : dsh plugin --profile web add dsh-vision-recognizer — no build scripts, no sharp approval (no native dependencies at all). - Configure from Settings → Plugins → Vision : pick a provider, enter an API key, override model / endpoint / token cap / timeout / marker. Saved changes take effect immediately, no restart. - 15+ providers, domestic and international : OpenAI, Anthropic Claude, Google Gemini, OpenRouter, Azure OpenAI, Ollama (local), plus Alibaba DashScope, QwenCloud (Intl), Zhipu GLM, Baidu Qianfan,…

在 GitHub 查看完整 README →
内容/媒体deepseek-harnessdsh-pluginvision

分类