Vision
视觉能力
56 个插件
- modlens1268
The first vision plugin for DeepSeek Harness, and the vision bridge for every text-only coding agent. Paste an image, get structured JSON evidence (OCR, layout, semantics). | 全网第一个 DeepSeek Harness 视觉插件,为 DeepSeek、GLM 等纯文本模型外挂视觉能力,粘贴图片即得结构化 JSON 证据(OCR、版面、语义)。
- agent-vision-toolkit829
为纯文本模型"看图“设计更好的视觉工具箱和技能,支持多图理解,图片问答,前端UI还原、GUI 自动化等,并可选无缝接入多个主流agent,直接识别粘贴图片| A vision toolkit and skill designed for text-only llms — image Q&A, long-screenshot OCR, frontend UI restoration, and GUI automation, with optional seamless integration for Codex, Claude Code, Pi, Oh My Pi, and OpenCode
- dsh-vision-toolkit324
让纯文本模型更好地做视觉任务的DeepSeek Harness插件:带意图的图片问答、长截图 OCR、UI 还原等|DeepSeek Harness-native integration for agent-vision-toolkit: image Q&A, long-screenshot OCR, UI restoration, grounding, pixel diff, Artifacts, and Web UI.
- dsh-vision-router47
Eyes for text-only DeepSeek Harness agents: built-in free vision chain (no key) + pixel-level vision tools (Q&A, grounding, crop, pixel diff, colors, OCR, SVG trace, cutout, screenshots). One-command install, no Python, image turns work like ordinary tool-calling turns.
- dsh-vision21
Near-native image understanding for DeepSeek Harness
- dsh-vision19
dsh 插件:给纯文本 DeepSeek 加视觉——view_image 工具桥接任意 OpenAI 兼容 VLM(默认智谱免费档,实测 4 厂商 10 模型)
- dsh-vision10
为 DeepSeek Harness 增加外挂识图模型:圆形鲸鱼按钮、发送图片识图自动回传、模型自主截图+识图工具、多协议自动适配、小白一键安装(未装 Node.js 自动下载)
- deepseek-harness-docker8
Community Docker and Kubernetes packaging for DeepSeek Harness (@deepseek-ai/dsh), with a hardened image, Compose stack, Helm chart, Web UI, and headless CLI.
- dsh-director-toolkit7
DSH Director Toolkit is a DeepSeek Harness plugin for 3D artists, technical designers, and creative coders. Paste a half-formed idea, a reference note, or a portfolio caption and get a compact direction pack for Blender, Three.js, Houdini, or C4D.
- image-vision7
暂无描述
- dsh-drop-to-path6
DSH 插件:图片与文件直达纯文本模型——图片保留原生附件体验,PDF/Office/压缩包/视频/音频显示为附件栏方块,点击发送时自动转为工作区路径,配合 dsh-vision-toolkit 粘贴即看图。A DSH plugin that delivers images AND files to text-only models as workspace paths: images keep the native attachment UI, other files show as square chips in the rail, paths append on send — pairs with dsh-vision-toolkit.
- dsh-vision-proxy6
DeepSeek Harness 插件:DeepSeek 大脑 + 自动识图。GUI 附加图片自动经 OpenAI 兼容 VLM 转译成文字后交给 DeepSeek 作答;支持百炼/智谱/OpenRouter 等任意 OpenAI 兼容端点(默认 qwen3.7-flash),无 key 自动探测本地 Ollama(图片不出本机);安装时有一问式确认
- dsh-image-subagent4
暂无描述
- dsh-plugin4
Upload images and files to your image host from DeepSeek Harness, powered by PicGo
- dsh-plugin-deepeye4
DeepEye vision plugin for DeepSeek Harness (DSH): image description, OCR, VQA, UI layout, and clipboard analysis.
- dsh-vision-opencode4
暂无描述
- dsh-vision-sidecar4
Hosted free vision sidecar for DeepSeek Harness with durable session evidence
- DeepSeek_Prism3
为纯文本 DeepSeek 模型提供按需识图的 Codex Skill(VEP/1 视觉证据包 + 多 Provider 降级)
- dsh-ernie-image3
暂无描述
- dsh-paddle-ocr3
暂无描述
- dsh-plugin-describe-image3
DeepSeek Harness plugin: describe_image — give a text-only model vision through an OpenAI-compatible VLM endpoint
- GrassVison3
给 DeepSeek 等纯文本大模型外挂图像理解能力的实现无感添加视觉能力。提供 OpenAI 兼容的 API,自动将图片请求交给视觉模型分析,再将结构化结果注入文本模型,使增强后的模型体验接近原生多模态。
- cli2
AtlasCloud CLI installers and release artifacts
- deepseek-harness-vision-plugin2
暂无描述
- dsh-cad-review2
Evidence-first ASCII DXF inspection and deterministic CAD rule review for DeepSeek Harness
- dsh-docling2
Native Docling document intelligence for DeepSeek Harness.
- dsh-image-bridge2
DSH 插件:让纯文本模型也能看图。Web 端直接粘贴图片即可发送,无需指定图片路径;模型自主调用视觉技能查看,多模态模型原生直通,零skill绑定。
- dsh-image-tools2
DSH bundle plugin: chat-image bridge + read_image deny + conversational image_recognize for text-only main models | 纯文本主模型识图桥接与识图工具
- dsh-multimodal2
给 DeepSeek 安装一双眼睛和一支画笔:会话里直接贴截图/图片,GLM 视觉模型先精确转写图片内容(报错信息、代码、界面逐字保留),然后 DeepSeek 继续处理你的问题——同一轮完成,全程无感;需要配图时,DeepSeek 自动调用文生图后端出图并显示在会话中。
- dsh-plugin-image-wallpaper2
自定义Deepseek Harness webUI主题
- dsh-read-image2
暂无描述
- dsh-sfversion2
SF视觉桥——给纯文本模型的 DeepSeek Harness 装上眼睛。
- dsh-tool-vision2
Vision model for DeepSeek Harness | DeepSeek Harness 外置视觉模型插件
- dsh-vision2
DeepSeek Harness 识图插件:为不具备原生识图能力的模型提供识图能力(阿里云百炼 qwen3.5-omni-plus,失败自动切换智谱 glm-4.6v-flash)。由 claude-vision-skill 移植适配。 | Vision tool for DeepSeek Harness
- dsh-vision2
DeepSeek Harness: vision
- dsh-vision-primitives2
Native interactive visual-reasoning plugin for DeepSeek Harness: precise pixel grounding (SOM grid / zoom / annotate / measure / diff / color / OCR) + MiMo V2.5 multimodal backend, zero external MCP servers.
- dsh-vision-provider2
Config-only DeepSeek Harness bundle for OpenAI-compatible vision models.
- dsh-xiapan-media2
Native vision, gpt-image-2 and Seedance plugins for DeepSeek Harness via Xiapan Cloud
- free-vision-skill2
Local‑only vision skill for macOS 本地化识图技能
- glm4v-vision-mcp2
GLM-4.6V 图像理解 MCP:识图/OCR/图表解析,原生接入 DeepSeek Harness(dsh-mcp-client),也兼容 Codex/Cline 等
- pi-mm-vision2
Synesthesia Encoder (通感编码器) — give any text-only LLM (DeepSeek, etc.) the ability to see images via structured spatial text encoding. A Pi agent extension.
- shadow-vision2
Open-source MCP vision server that gives text-only LLMs and AI agents image understanding, OCR, visual analysis, UI inspection, and multimodal capabilities.
- dsh-mimo-vision-hint1
DSH plugin: dispatch image-recognition tasks to an opencode-go mimo-v2.5 subagent via system-prompt injection
- @bujue3184/dsh-tool-vision0
Vision tools (analyze_image, locate_element) for DeepSeek Harness: local ollama, DSH subagent (qwen-vl etc), or OpenAI-compatible HTTP endpoints
- @evan-williams/vision-tool0
Conversation-style UI/UX visual review plugin for DeepSeek Harness: vision_review / vision_ask tools backed by any OpenAI-compatible multimodal endpoint (e.g. agnes-2.5-flash)
- @liu__min/dsh-vision-bridge0
DSH 视觉桥接插件:让无视觉能力的主模型看图(会话收图 + 自动转文字 + view_image 工具)
- @niyongsheng/free-vision-skill0
DSH-Plugin for DeepSeek-Harness: fully-local image understanding & OCR powered by macOS Vision Framework
- dsh-multimodal-bridge0
DeepSeek Harness plugin bundle: qwen_vision (Qwen-VL image understanding) and qwen_generate (Qwen-Image text-to-image) tools for text-only models
- dsh-nanobananapro0
Generate images and videos in DeepSeek Harness through the NanoBananaPro API
- dsh-plugin-mm-vision0
mm-vision (通感编码器) for DeepSeek Harness — give any text-only LLM the ability to see images via structured spatial text encoding. Registers the mm_vision tool.
- dsh-plugin-vision-toolkit0
Vision toolkit for DeepSeek Harness -- glance, ground, detect, crop CLI tools for text-only agents to understand images
- dsh-seedance20
Generate images and Seedance videos in DeepSeek Harness through the Seedance 2 AI API
- dsh-sight0
Plug-in vision for text-only DeepSeek Harness (dsh) models: a `vision` tool with built-in cheap/free VLM presets, multi-image batch analysis, paste-to-hint image admission, and a web settings page with hot-reload.
- dsh-usage-dashboard-plus0
DeepSeek Harness (dsh) web plugin: API balance + today's spend sidebar widget, /api/dsh-usage route, external vision-call accounting (JSONL usage log), plus a full per-session usage dashboard (call log, per-model stats, cache rate, CSV export) via the usa
- dsh-vision-tools0
DeepSeek Harness 视觉能力全家桶:vision_understand 工具(OpenAI 兼容视觉 API,默认免费智谱 GLM-4V-Flash)+ 粘贴/拖拽/按钮三入口识图
- dsh-yali-image-generator0
DeepSeek-Harness 图像生成插件。申请 Yali AI API Key:https://api.yaliai.com/