All Plugins

29 plugins page 1 of 2

Results

  • Listed Bundle verified

    Free vision for text-only agents: built-in keyless vision chain plus pixel tools (Q&A, grounding, crop, pixel diff, colors, OCR, SVG trace, cutout, screenshots); paste an image to use it.

    dsh plugin --profile default add github:ysr666/dsh-vision-router
  • Listed Bundle verified

    Image "reading" for text-only models: downscale + reduce color depth + structure/color fingerprints into text grids fed back to the conversation, letting the model zoom, sample and OCR autonomously like a multimodal model; fully local with zero external model dependency, ships an image-reading methodology skill and optional PaddleOCR.

    dsh plugin --profile default add github:jing-hy/picturereader
  • Listed Bundle verified

    External vision plugin for DeepSeek Harness: whale-button config panel, image recognition with auto-reply, and agent screenshot/recognize tools.

    dsh plugin --profile default add github:linenxi-ctrl/dsh-vision
  • Listed Bundle verified

    DeepSeek brain + automatic image transcription: attach images in the GUI and each one is transcribed to text via any OpenAI-compatible VLM before reaching the text-only DeepSeek — a keyed fast path (default qwen3.7-flash; DashScope/Zhipu/OpenRouter or any OpenAI-compatible endpoint) with your own key, or local Ollama auto-detected with zero config.

    dsh plugin --profile default add github:Flyvhidbwo/dsh-vision-proxy
  • Listed Bundle verified

    Free vision bridge and image generation for text-only models: paste-image reading, GLM-4V-Flash and Gemini engine failover, ModLens-style structured evidence, and a seeded free vision model route.

    dsh plugin --profile default add github:MJorgin/dsh-media-skills
  • Listed Bundle verified

    MCP bridge to gemini.google.com: vision analysis of images and videos, Imagen image and Veo video generation, and conversation management using the logged-in browser session with no API key.

    dsh plugin --profile default add github:ConsoleSun/Gemini-Eyes
  • Listed Bundle verified

    Local OCR for attached images via the built-in Windows engine (Windows.Media.Ocr): only the recognized text is sent to the model, never the image bytes; vision passthrough is opt-in.

    dsh plugin --profile default add github:maxwell-feng/dsh-windows-ocr
  • Listed Bundle verified

    Automatically generates and displays images in the DSH chat via API channels or local CLIs (supports mmx / codex / agy).

    dsh plugin --profile default add github:corrinehu/dsh-chat-imagine
  • Listed Bundle verified

    MiniMax-powered multimodal plugin: real-time voice call mode (streaming conversation, floating dock UI), voice mode and mic voice input, plus image/video/music/speech generation and vision inspection tools.

    dsh plugin --profile default add github:AtropinolTT/dsh-guide-dog
  • Listed Bundle verified

    Vision MCP and DSH bundle for text-only DeepSeek: analyze_image, analyze_clipboard, compare_images and vision_status tools, a visual settings page, free GLM-4.6V-Flash by default, result caching and rate-limit tolerance; keys stay out of logs.

    dsh plugin --profile default add github:GOU-GEE/deepseek-vision#path:/plugins/dsh-plugin-deepseek-vision
  • Listed Bundle verified

    Native LLM-provider vision bridge: images pasted in the chat are described by a vision model (Qwen3-VL via pi-ai/llama.cpp) and the text description is fed to text-only DeepSeek for the reply — image admission, routing and compaction all run through harness-native mechanisms, with an LRU description cache and 503 retry.

    dsh plugin --profile default add github:Einskyle/dsh-llm-vision-bridge
  • Listed Bundle verified

    Composer-attached images are transcribed to text by an OpenAI-compatible vision model before reaching text-only DeepSeek models.

    dsh plugin --profile default add github:ximengxiaolan/dsh-vision-bridge
  • Listed Bundle verified

    Local OCR for attached images via Tesseract: only the recognized text is sent to the model, never the image bytes; vision passthrough is opt-in.

    dsh plugin --profile default add github:maxwell-feng/dsh-tesseract-ocr
  • Listed Bundle verified

    Full vision-capability bundle for DeepSeek Harness: a vision_understand tool (OpenAI-compatible vision APIs, free Zhipu GLM-4V-Flash by default) plus paste/drag-and-drop/button entry points for image recognition.

    dsh plugin --profile default add github:moon09300731/dsh-vision-tools
  • Listed Bundle verified

    Free vision bridge for text-only models: image understanding, OCR, UI and debug analysis via free-tier providers (Qwen3-VL-Flash, Doubao, DeepSeek-OCR) with a settings GUI.

    dsh plugin --profile default add github:FuzzySoul/dsh-free-vision
  • Listed Bundle verified

    Universal image generation for DeepSeek Harness: auto-discovers image models from any OpenAI-compatible endpoint (SenseNova, StepFun, Agnes, Qwen, Flux, SD, Imagen and more), with agent tools and REST API.

    dsh plugin --profile default add github:xiaozhe7772222/dsh-draw-router
  • Listed Bundle verified

    Adds an image_describe tool backed by the DashScope OpenAI-compatible vision API; on text-only sessions, pasted images are stored as local paths for the model and rendered inline in the chat transcript.

    dsh plugin --profile default add github:wanshichenguang/dsh-tool-vision
  • Listed Bundle verified

    314857493/dsh-vision#vision-route: Registers a `deepseek-vision` provider route: the Web GUI accepts pasted images and transcribes them to text via the free Zhipu GLM vision API before delegating to the DeepSeek adapter. 314857493/dsh-vision#vision-tool: Model-facing `vision` tool for DeepSeek Harness: describe and OCR image files by calling the free Zhipu GLM vision API directly (glm-4v-flash fallback chain), no external CLI required.

    dsh plugin --profile default add dsh-vision-proxy-route
  • Listed Bundle verified

    Reuses DeepSeek web's built-in vision mode for text-only models: the deepseek_vision tool drives the local deepseek-vision-cli browser automation (manual login helper, deep-think enabled, auto-closes browser) and returns image descriptions as text.

    dsh plugin --profile default add github:Cheng-cheng9669/dsh-deepseek-vision
  • Listed Bundle verified

    On-demand vision for text-only DeepSeek models: upload images, and the model calls a view_image tool backed by any OpenAI-compatible vision endpoint (Qwen/DashScope by default).

    dsh plugin --profile default add github:Leeminjing/dsh-eyes