All Plugins
Results
-
Listed
Bundle verified
Vision bridge for text-only models: paste an image, get structured JSON evidence (OCR, layout, semantics).
dsh plugin --profile default add github:liustack/modlens -
Listed
Bundle verified
Free vision for text-only agents: built-in keyless vision chain plus pixel tools (Q&A, grounding, crop, pixel diff, colors, OCR, SVG trace, cutout, screenshots); paste an image to use it.
dsh plugin --profile default add github:ysr666/dsh-vision-router -
Listed
Bundle verified
Vision tasks for text-only models: intent-aware image Q&A, long-screenshot OCR, UI reproduction, grounding, and pixel diff.
dsh plugin --profile default add github:Anionex/dsh-vision-toolkit -
Listed
Bundle verified
Image "reading" for text-only models: downscale + reduce color depth + structure/color fingerprints into text grids fed back to the conversation, letting the model zoom, sample and OCR autonomously like a multimodal model; fully local with zero external model dependency, ships an image-reading methodology skill and optional PaddleOCR.
dsh plugin --profile default add github:jing-hy/picturereader -
Listed
Bundle verified
External vision plugin for DeepSeek Harness: whale-button config panel, image recognition with auto-reply, and agent screenshot/recognize tools.
dsh plugin --profile default add github:linenxi-ctrl/dsh-vision -
Listed
Bundle verified
DeepSeek brain + automatic image transcription: attach images in the GUI and each one is transcribed to text via any OpenAI-compatible VLM before reaching the text-only DeepSeek — a keyed fast path (default qwen3.7-flash; DashScope/Zhipu/OpenRouter or any OpenAI-compatible endpoint) with your own key, or local Ollama auto-detected with zero config.
dsh plugin --profile default add github:Flyvhidbwo/dsh-vision-proxy -
Listed
Bundle verified
Gives text-only models vision: forwards user images to an OpenAI-compatible vision model and shows the descriptions in a Web UI right panel.
dsh plugin --profile default add github:jyh20030112/dsh-visual-plugin -
Listed
Bundle verified
Free vision bridge and image generation for text-only models: paste-image reading, GLM-4V-Flash and Gemini engine failover, ModLens-style structured evidence, and a seeded free vision model route.
dsh plugin --profile default add github:MJorgin/dsh-media-skills -
Listed
Bundle verified
MCP bridge to gemini.google.com: vision analysis of images and videos, Imagen image and Veo video generation, and conversation management using the logged-in browser session with no API key.
dsh plugin --profile default add github:ConsoleSun/Gemini-Eyes -
Listed
Bundle verified
Vision for text-only DeepSeek via Doubao Web by default (zero-cost, no API key — drives your logged-in Chrome through a Windows CDP bridge), with Antigravity IDE quota (flash/pro) or Gemini fallback; auto detail escalation, vision evidence memory with compaction rehydration, content-hash cache, and a bilingual client panel.
dsh plugin --profile default add github:54xkeee/dsh-vision -
Listed
Bundle verified
A vision-language gateway provider route: pasted images are described by a configurable VL model (Qwen-VL by default) before the DeepSeek wire.
dsh plugin --profile default add github:siegfly/dsh-deepseek-vision -
Listed
Bundle verified
Local OCR for attached images via the built-in Windows engine (Windows.Media.Ocr): only the recognized text is sent to the model, never the image bytes; vision passthrough is opt-in.
dsh plugin --profile default add github:maxwell-feng/dsh-windows-ocr -
Listed
Bundle verified
Automatically generates and displays images in the DSH chat via API channels or local CLIs (supports mmx / codex / agy).
dsh plugin --profile default add github:corrinehu/dsh-chat-imagine -
Listed
Bundle verified
AI image generation for the DSH Web GUI: text-to-image and image-to-image through a configurable OpenAI-compatible endpoint (gpt-image-2 / gpt-image-1 / dall-e-3), with an api_url/api_key settings card and a sidebar split-pane generation studio.
dsh plugin --profile default add github:dickpy/dsh-imagegen -
Listed
Bundle verified
MiniMax-powered multimodal plugin: real-time voice call mode (streaming conversation, floating dock UI), voice mode and mic voice input, plus image/video/music/speech generation and vision inspection tools.
dsh plugin --profile default add github:AtropinolTT/dsh-guide-dog -
Listed
Bundle verified
Vision MCP and DSH bundle for text-only DeepSeek: analyze_image, analyze_clipboard, compare_images and vision_status tools, a visual settings page, free GLM-4.6V-Flash by default, result caching and rate-limit tolerance; keys stay out of logs.
dsh plugin --profile default add github:GOU-GEE/deepseek-vision#path:/plugins/dsh-plugin-deepseek-vision -
Listed
Bundle verified
Fully-local image understanding & OCR via macOS Vision Framework: `ocr_image` (text, table layout + coordinates) and `view_image` (scene, faces, QR) — paste multiple images into the web input box or pass path/URL/base64; images never leave your Mac.
dsh plugin --profile default add github:niyongsheng/free-vision-skill -
Listed
Bundle verified
Advertise image paste on text-only DeepSeek routes, describe attachments with a vision sidecar, and leave native vision models untouched.
dsh plugin --profile default add github:shinjiyu/dsh-plugin-multimodal -
Listed
Bundle verified
Native LLM-provider vision bridge: images pasted in the chat are described by a vision model (Qwen3-VL via pi-ai/llama.cpp) and the text description is fed to text-only DeepSeek for the reply — image admission, routing and compaction all run through harness-native mechanisms, with an LRU description cache and 503 retry.
dsh plugin --profile default add github:Einskyle/dsh-llm-vision-bridge -
Listed
Bundle verified
Composer-attached images are transcribed to text by an OpenAI-compatible vision model before reaching text-only DeepSeek models.
dsh plugin --profile default add github:ximengxiaolan/dsh-vision-bridge