DeepSeek Agent
插件市场界面与桌面dsh-vision-toolkit
DEEPSEEK HARNESS PLUGIN

dsh-vision-toolkit

[dsh]为纯文本模型设计更强大的视觉工具箱:一行安装使用、粘贴图片直接识别、多张图片问答、截图到前端UI 还原等|DeepSeek Harness-native integration for agent-vision-toolkit: image Q&A, long-screenshot OCR, UI restoration, grounding, pixel diff, Artifacts, and Web UI.

插件介绍

[dsh]为纯文本模型设计更强大的视觉工具箱:一行安装使用、粘贴图片直接识别、多张图片问答、截图到前端UI 还原等|DeepSeek Harness-native integration for agent-vision-toolkit: image Q&A, long-screenshot OCR, UI restoration, grounding, pixel diff, Artifacts, and Web UI.

agent-skillsagent-vision-toolkitcomputer-visiondeepseekdeepseek-harnessdshdsh-plugingui-automationocrpluginpythonscreenshot-testingtext-only-llmtypescriptui-restorationvision-language-modelvision-tools

项目详情摘要

DSH Vision Toolkit A more powerful vision toolkit—give text-only models in DeepSeek Harness eyes: image Q&A, long-screenshot OCR, UI restoration, and GUI visual tasks in one toolkit and Skill.** 🚀 Paste an image and ask directly | Install with one command | Broad use cases Highlights | Quick start | Toolbox | Configuration and limits | Troubleshooting | Community 🌐 **English** | 中文 🏆 This project is the first comprehensive vision-tool plugin in the DeepSeek Harness ecosystem: it was initiated before internal beta and built during the beta with reference to `agent-vision-toolkit`. **Original work:** The system and division of responsibilities behind these visual tools, together with the `vision-skills` Skill, were personally created and continuously refined by the author through long-term real-world use and repeated iteration. If this project helps you or gives you some inspiration, feel free to star 🌟 & fork. Highlights **Paste an image and ask directly.** In DSH Web, pasting an image switches the text-only model to its `(Vision Toolkit)` variant automatically — no manual path copying or model changes. Native thumbnails, session history, and workspace paths stay intact; Web can preview artifacts. **One command to install.** After installation, configure a vision provider in **Settings → Vision Toolkit** and start using the tools. **Not just a caption — the content that matters.** The model does not produce a generic description; it extracts evidence around the current task, such as “Where is the error?” or “Where is the button?”. **A battle-tested visual-task methodology.** The bundled Skill tells the agent what to look at for different visual tasks, which tool to choose, how to proceed, and how to verify the result. `agent-vision-toolkit` gives an agent more than image

摘自项目公开 README,可能随上游仓库更新。

安装方法

建议先在测试 Profile 中安装,并检查权限、安装脚本和依赖。

npx -p @deepseek-ai/dsh dsh plugin --profile web add github:Anionex/dsh-vision-toolkit

使用前检查

  • 确认项目符合 DSH bundle 规范,而不只是相关仓库。
  • 阅读许可证和安装脚本,检查网络、文件及执行权限。
  • 备份配置,并确保插件能够安全卸载或回滚。