MJorgin/dsh-media-skills

MJorgin★ 8PythonLast synced: 2026-08-17

Open on GitHub

Free vision & image generation for DeepSeek Harness — paste an image into any chat, even text-only sessions. GLM-4V-Flash / Qwen3-VL / Gemini failover chain, ModLens-style structured evidence, Kolors generation. 免费读图·生图 · 三引擎容错 · 无 Key 入库

README excerpt

🎨 dsh-media-skills Give DeepSeek Harness eyes — and a brush. Read images in any chat, generate new ones, all with free models. DeepSeek Harness is brilliant at reasoning — but a text-only model can't see the image you just dragged into the chat. This bundle fixes that with two free skills , a free vision model route , and a vision engine failover chain : - 📎 Paste to read — paste, drag, or pick an image in any session; the free vision model turns it into text your current model understands. (Powered by the DeepSeek Harness core auto-description path — see docs/HARNESS PATCH EN.md; this bundle contributes the vision model route and the skill it relies on.) - 👁️ vision-review — analyze images and screenshots, catch UI visual bugs, detect watermarks, turn images into text. - 🎨 media-tools — generate illustrations, avatars, backgrounds and banners with a free, watermark-free model. - 🔀 Engine failover — GLM-4V-Flash → SiliconFlow Qwen3-VL → Google Gemini (AI Studio) → any OpenAI-compatible endpoint, with ModLens-style structured evidence output. No hardcoded keys, no paid API, no file saving, no session switching. Why · Quick start · See it in action · Usage · Keys & privacy · FAQ…

View full README on GitHub →
Media / Contentagent-skillsdeepseek-harnessdsh-pluginimage-generationskillvision

Category