jin123-alpha/dsh-ext-vision-proxy

jin123-alpha★ 1TypeScript最后同步: 2026-08-19

在 GitHub 打开

External vision proxy extension for DeepSeek Harness, enabling text-only models to analyze and understand images via OpenAI-compatible vision APIs.

README 摘要

dsh-ext-vision-proxy English 简体中文 External vision proxy plugin for DeepSeek Harness, enabling text-only LLMs (such as DeepSeek-V3 / DeepSeek-R1) to inspect, analyze, and understand images via OpenAI-compatible vision model APIs. 📸 Screenshots 1. Settings Panel (Configuration & Provider Management) 2. Chat Composer Switch (Pill-shaped Toggle & Model Selector) 🌟 Key Features - 🖼️ Multimodal Power for Text-Only Models : Allows text-only LLMs to autonomously invoke the vision describe tool to inspect and answer questions about images, screenshots, diagrams, and attachments. - 🔘 Native Composer Switch : Integrates seamlessly into the chat input bar ( conversation.input.left ) with a pill-shaped toggle switch and an interactive model dropdown selector. - ⚙️ Dedicated Settings & Connectivity Testing : Provides an intuitive configuration panel in the Settings page supporting custom Base URLs, API Keys, vision model filtering, and one-click connectivity testing. - 🎛️ 4-State Session Matrix : 1. Multimodal LLM + Switch ON : vision describe tool is visible. Model can choose native or delegated recognition. Prompts confirmation dialog upon image attachment. 2. Multimodal LLM + Switch OFF …

在 GitHub 查看完整 README →
终端/TUIdshdsh-pluginvision

分类