Application-level Vision-Language-Model (VLM) analyzer for DeepSeek Harness: analyze_image tool with primary/backup OpenAI-compatible endpoints, automatic failover, and an auto-saving web settings page.
视觉
$ npx -p @deepseek-ai/dsh dsh plugin --profile web add github:huashenglian/dsh-her-eyes
Give text-only models eyes: an analyze_image tool for DeepSeek Harness, backed by free Chinese vision APIs (GLM-4V-Flash / Qwen-VL) or any OpenAI-compatible vision endpoint. 给纯文本模型装上眼睛。
Agent 协作 · integrations · 视觉
$ npx -p @deepseek-ai/dsh dsh plugin --profile web add github:gmleong/dsh-image-bridge
Let image-incapable LLM models receive image messages in DeepSeek Harness (dsh): replace image blocks with text placeholders + local paths, then describe via a vision script (qwen)
Give DeepSeek Harness agents the ability to read images directly: a model-facing read_image tool that answers questions about an image through any OpenAI-compatible vision endpoint.
Agent 协作 · 视觉
$ npx -p @deepseek-ai/dsh dsh plugin --profile web add github:zcXie777/dsh-image-reader
Jina AI tools for DeepSeek Harness: search (incl. arxiv/ssrn academic domains), read, screenshot, embed, rerank, classify, pdf, expand, datetime, primer — with a settings-page API key UI.
finance · 界面 · 视觉
$ npx -p @deepseek-ai/dsh dsh plugin --profile web add github:minatoAI/jina-dsh-plugin
Kimi WebBridge for DeepSeek Harness: drive the user's real browser (navigate, click, fill, snapshot, screenshot, evaluate, network, upload, PDF) through the local Kimi WebBridge daemon
效率 · 视觉
$ npx -p @deepseek-ai/dsh dsh plugin --profile web add github:MicroHEROX/dsh-Kimi-WebBridge
KoboldCpp for DeepSeek Harness - a tool plugin that lets the harness online model hand repetitive text and vision (OCR) labor to a local KoboldCpp (llama.cpp) server.
Qwen-MM-Plugins integration for DeepSeek Harness: 12 MCP multimodal tools (vision/OCR/grounding/ASR/AV), a Web settings page (paste a Qwen API Key), bundled skills, and a one-command installer.
DSH plugin: free web search tool (DuckDuckGo via ddg-kit with system-proxy support, automatic Bing fallback) + front-loaded vision understanding tool (vision_understand: understand images via a vision-capable model configured in DSH when the conversation model has no image modality) + composer-level paste-image capture (paste/drop an image in the conversation input → vision model → description, bypassing DSH's image gate on text-only models). Official bundle plugin, install: dsh plugin --profile web add github:LingyeSoul/dsh-rider#main。DSH 插件:免费网络搜索(duckduckgo_search 工具)+ 前置视觉理解(vision_understand 工具)+ 对话输入框粘贴图片捕获(粘贴/拖入图片 → 视觉模型 → 描述,绕开 DSH 对纯文本模型的图片拦截)。
视觉
$ npx -p @deepseek-ai/dsh dsh plugin --profile web add github:LingyeSoul/dsh-rider
Agent-to-user image display for the DeepSeek Harness WebUI: a show_image tool whose results render as an inline image card in the conversation (via embedded presentation data), never entering model history — safe for text-only LLMs.
Customizable skinning tool for the DeepSeek Harness Web UI — presets, image wallpapers, translucency, and accent colors, all persisted in user settings