54xkeee/dsh-youreyes

54xkeee★ 1TypeScriptLast synced: 2026-08-15

Open on GitHub

Eyes for text-only DeepSeek on DeepSeek Harness: model-invokable vision tool + wrapper adapters + general VLM channels (OpenAI-compatible / Gemini / local Ollama)

README excerpt

👁️ dsh-youreyes Eyes for text-only DeepSeek. Paste images, screenshots, or file paths into DeepSeek Harness — and the model can finally "see" and answer image-related questions. DeepSeek stays the brain; vision is just the eyes. 简体中文 =20" / TL;DR : DeepSeek can't see images? Install this — paste, recognize, answer. Three steps. ✨ Why you'll love it Pain point dsh-youreyes solution DeepSeek is text-only; pasting an image is rejected Wrapper adapters claim image input; images become text placeholders automatically Other vision plugins lock you into one vendor Antigravity (default) + any OpenAI-compatible endpoint + Gemini + local Ollama — your key just works Paying per image recognition Antigravity IDE quota by default (flash/pro tiers) when the IDE is running; free local Ollama otherwise Setup is a chore with registrations everywhere Local Ollama auto-detection, zero config ; one line for a free Gemini key The model "forgets" what it saw Vision evidence memory : results persist in the session, reused across turns, restored after compaction Paying to re-recognize the same image Content-hash cache : same image + same question = recognized once per process Complex images get shallow a…

View full README on GitHub →
Media / Contentdeepseekdshdsh-plugingeminiimage-recognitionmultimodalollamavision

Category