All Plugins
Results
-
Listed
Bundle verified
Free vision for text-only agents: built-in keyless vision chain plus pixel tools (Q&A, grounding, crop, pixel diff, colors, OCR, SVG trace, cutout, screenshots); paste an image to use it.
dsh plugin --profile default add github:ysr666/dsh-vision-router -
Listed
Bundle verified
Image "reading" for text-only models: downscale + reduce color depth + structure/color fingerprints into text grids fed back to the conversation, letting the model zoom, sample and OCR autonomously like a multimodal model; fully local with zero external model dependency, ships an image-reading methodology skill and optional PaddleOCR.
dsh plugin --profile default add github:jing-hy/picturereader -
Listed
Bundle verified
External vision plugin for DeepSeek Harness: whale-button config panel, image recognition with auto-reply, and agent screenshot/recognize tools.
dsh plugin --profile default add github:linenxi-ctrl/dsh-vision -
Listed
Bundle verified
DeepSeek brain + automatic image transcription: attach images in the GUI and each one is transcribed to text via any OpenAI-compatible VLM before reaching the text-only DeepSeek — a keyed fast path (default qwen3.7-flash; DashScope/Zhipu/OpenRouter or any OpenAI-compatible endpoint) with your own key, or local Ollama auto-detected with zero config.
dsh plugin --profile default add github:Flyvhidbwo/dsh-vision-proxy -
Listed
Bundle verified
Free vision bridge and image generation for text-only models: paste-image reading, GLM-4V-Flash and Gemini engine failover, ModLens-style structured evidence, and a seeded free vision model route.
dsh plugin --profile default add github:MJorgin/dsh-media-skills -
Listed
Bundle verified
MCP bridge to gemini.google.com: vision analysis of images and videos, Imagen image and Veo video generation, and conversation management using the logged-in browser session with no API key.
dsh plugin --profile default add github:ConsoleSun/Gemini-Eyes -
Listed
Bundle verified
Local OCR for attached images via the built-in Windows engine (Windows.Media.Ocr): only the recognized text is sent to the model, never the image bytes; vision passthrough is opt-in.
dsh plugin --profile default add github:maxwell-feng/dsh-windows-ocr -
Listed
Bundle verified
Automatically generates and displays images in the DSH chat via API channels or local CLIs (supports mmx / codex / agy).
dsh plugin --profile default add github:corrinehu/dsh-chat-imagine -
Listed
Bundle verified
MiniMax-powered multimodal plugin: real-time voice call mode (streaming conversation, floating dock UI), voice mode and mic voice input, plus image/video/music/speech generation and vision inspection tools.
dsh plugin --profile default add github:AtropinolTT/dsh-guide-dog -
Listed
Bundle verified
Vision MCP and DSH bundle for text-only DeepSeek: analyze_image, analyze_clipboard, compare_images and vision_status tools, a visual settings page, free GLM-4.6V-Flash by default, result caching and rate-limit tolerance; keys stay out of logs.
dsh plugin --profile default add github:GOU-GEE/deepseek-vision#path:/plugins/dsh-plugin-deepseek-vision -
Listed
Bundle verified
Native LLM-provider vision bridge: images pasted in the chat are described by a vision model (Qwen3-VL via pi-ai/llama.cpp) and the text description is fed to text-only DeepSeek for the reply — image admission, routing and compaction all run through harness-native mechanisms, with an LRU description cache and 503 retry.
dsh plugin --profile default add github:Einskyle/dsh-llm-vision-bridge -
Listed
Bundle verified
Composer-attached images are transcribed to text by an OpenAI-compatible vision model before reaching text-only DeepSeek models.
dsh plugin --profile default add github:ximengxiaolan/dsh-vision-bridge -
Listed
Bundle verified
Local OCR for attached images via Tesseract: only the recognized text is sent to the model, never the image bytes; vision passthrough is opt-in.
dsh plugin --profile default add github:maxwell-feng/dsh-tesseract-ocr -
Listed
Bundle verified
Full vision-capability bundle for DeepSeek Harness: a vision_understand tool (OpenAI-compatible vision APIs, free Zhipu GLM-4V-Flash by default) plus paste/drag-and-drop/button entry points for image recognition.
dsh plugin --profile default add github:moon09300731/dsh-vision-tools -
Listed
Bundle verified
Free vision bridge for text-only models: image understanding, OCR, UI and debug analysis via free-tier providers (Qwen3-VL-Flash, Doubao, DeepSeek-OCR) with a settings GUI.
dsh plugin --profile default add github:FuzzySoul/dsh-free-vision -
Listed
Bundle verified
Universal image generation for DeepSeek Harness: auto-discovers image models from any OpenAI-compatible endpoint (SenseNova, StepFun, Agnes, Qwen, Flux, SD, Imagen and more), with agent tools and REST API.
dsh plugin --profile default add github:xiaozhe7772222/dsh-draw-router -
Listed
Bundle verified
Adds an image_describe tool backed by the DashScope OpenAI-compatible vision API; on text-only sessions, pasted images are stored as local paths for the model and rendered inline in the chat transcript.
dsh plugin --profile default add github:wanshichenguang/dsh-tool-vision -
Listed
Bundle verified
TZHR-invest/dsh-plugins#dsh-lan-gateway: Complete LAN/remote access for the Web UI: 0.0.0.0 binding, crypto.randomUUID polyfill, token gate (401 login page + WebSocket interception, loopback exempt), privileged-fence and settings-persistence exemptions, and an idempotent installer with upgrade recovery. TZHR-invest/dsh-plugins#dsh-mobile-ui: Mobile UI for the Web GUI on narrow screens: full-width responsive layout, overlay session drawer, 44px touch targets, safe-area support, and reading enhancements with zero desktop impact. TZHR-invest/dsh-plugins#dsh-vision-tool: Agent-callable vision tool that describes local images via any OpenAI-compatible vision endpoint you configure, with an optional multi-model cross-check and no built-in keys. TZHR-invest/dsh-plugins#dsh-web-search-metaso: Metaso (秘塔AI搜索) search & reader providers for the web seam: web_search returns page summaries, web_fetch reads full-page markdown, multi-scope search (webpage/document/paper/image/video/podcast).
dsh plugin --profile default add github:TZHR-invest/dsh-plugins#path:/packages/dsh-lan-access -
Listed
Bundle verified
314857493/dsh-vision#vision-route: Registers a `deepseek-vision` provider route: the Web GUI accepts pasted images and transcribes them to text via the free Zhipu GLM vision API before delegating to the DeepSeek adapter. 314857493/dsh-vision#vision-tool: Model-facing `vision` tool for DeepSeek Harness: describe and OCR image files by calling the free Zhipu GLM vision API directly (glm-4v-flash fallback chain), no external CLI required.
dsh plugin --profile default add dsh-vision-proxy-route -
Listed
Bundle verified
Reuses DeepSeek web's built-in vision mode for text-only models: the deepseek_vision tool drives the local deepseek-vision-cli browser automation (manual login helper, deep-think enabled, auto-closes browser) and returns image descriptions as text.
dsh plugin --profile default add github:Cheng-cheng9669/dsh-deepseek-vision