catalog / media-vision
设计、媒体与视觉
我们精选列表中所有「设计、媒体与视觉」插件,与 GitHub 同步。
同步自我们的 GitHub 列表 · 目录更新于 2026-08-17
93 个插件
modlens
作者 liustack
The first vision plugin for DeepSeek Harness, and the vision bridge for every text-only coding agent. Paste an image, get structured JSON evidence (OCR, layout, semantics). | 全网最强 DeepSeek Harness 外挂视觉插件,为 DeepSeek、GLM 等纯文本模型外挂视觉能力,粘贴图片即得结构化 JSON 证据(OCR、版面、语义)。
dsh-vision-toolkit
作者 anionex
让纯文本模型更好地做视觉任务的DeepSeek Harness插件:带意图的图片问答、长截图 OCR、UI 还原等|DeepSeek Harness-native integration for agent-vision-toolkit: image Q&A, long-screenshot OCR, UI restoration, grounding, pixel diff, Artifacts, and Web UI.
dsh-vision-router
作者 ysr666
Eyes for text-only DeepSeek Harness agents: built-in fr…
dsh-openpencil
作者 zseven-w
连接 DeepSeek Harness 与 OpenPencil,让智能体创建、编辑、预览和验证可交互的多页面设计画布。
dsh-vision
作者 oil-oil
Near-native image understanding for DeepSeek Harness
dsh-vision
作者 william-jin-cmu
dsh 插件:给纯文本 DeepSeek 加视觉——view_image 工具桥接任意 OpenAI 兼容 VLM(默认智谱免费档,实测 4 厂商 10 模型)
dsh-vision
作者 linenxi-ctrl
为 DeepSeek Harness 增加外挂识图模型:圆形鲸鱼按钮、发送图片识图自动回传、模型自主截图+识图工具、多协议自动适配、小白一键安装(未装 Node.js 自动下载)
deepseek-visionary
作者 xlight
使用 DeepSeek 官方多模态视觉模型让你的 Agent 不再眼瞎(支持 DSH、Zed、OpenCode、Codex、Claude Code、Cursor、Claude Desktop)
dsh-vision-proxy
作者 flyvhidbwo
DeepSeek Harness 插件:DeepSeek 大脑 + 自动识图。GUI 附加图片自动经 OpenAI 兼容 VLM 转译成文字后交给 DeepSeek 作答;支持百炼/智谱/OpenRouter 等任意 OpenAI 兼容端点(默认 qwen3.7-flash),无 key 自动探测本地 Ollama(图片不出本机);安装时有一问式确认
dsh-plugin-aigc-canvas
作者 huanlinoto
provider-agnostic AIGC HTTP 桥 + 无限画布 + ffmpeg 后处理,13 个工具含画布连边/reroll/媒体编辑 | Provider-agnostic AIGC HTTP bridge + infinite canvas + ffmpeg post-processing; 13 tools incl. canvas linking/reroll/media-edit
dsh-visual-plugin
作者 jyh20030112
Dsh-visual-plugin.Give your text-only model eyes: forward user images to any OpenAI-compatible vision model and see the results in a Web UI right panel
image-vision
作者 wangyang10
视觉插件/技能:让纯文本模型调 OpenAI 兼容识图 API 看图(描述/问答/OCR/多图对比),DSH 提供 vision_query 工具 + /image-vision 斜杠命令
dsh-media-skills
作者 akqwpeter-prog
Free vision & image generation for DeepSeek Harness — paste an image into any chat, even text-only sessions. GLM-4V-Flash / Qwen3-VL / Gemini failover chain, ModLens-style structured evidence, Kolors generation. 免费读图·生图 · 三引擎容错 · 无 Key 入库
gemini-eyes
作者 consolesun
MCP bridge to gemini.google.com: vision analysis of images and videos, Imagen image and Veo video generation, and conversation management using the logged-in browser session with no API key.
agent_extensions
作者 dddfxyqiming
Agent Skills & DeepSeek Harness (DSH) 扩展库:通用智能体技能(General_skills)+ DSH 标准插件(dsh-plugin),开箱即用的 AI Agent 能力增强集合。
dsh-vision-sidecar
作者 121103qwq
Hosted free vision sidecar for DeepSeek Harness with durable session evidence
dsh-multimodal
作者 mc5lan
给 DeepSeek 安装一双眼睛和一支画笔:会话里直接贴截图/图片,GLM 视觉模型先精确转写图片内容(报错信息、代码、界面逐字保留),然后 DeepSeek 继续处理你的问题——同一轮完成,全程无感;需要配图时,DeepSeek 自动调用文生图后端出图并显示在会话中。
dsh-vision
作者 54xkeee
Vision for DeepSeek Harness: Doubao Web by default (zero-cost, no API key), Antigravity IDE quota (flash/pro), any IDE CLI, Gemini — auto detail escalation, evidence memory
dsh-plugin-deepeye
作者 favio8
DeepEye vision plugin for DeepSeek Harness (DSH): image description, OCR, VQA, UI layout, and clipboard analysis.
deepseek-vision
作者 gou-gee
Vision MCP and DSH bundle for text-only DeepSeek: analyze_image, analyze_clipboard, compare_images and vision_status tools, a visual settings page, free GLM-4.6V-Flash by default, result caching and rate-limit tolerance; keys stay out of logs.
dsh-her-eyes
作者 huashenglian
一个可以让ai自动调用VLM(多模态模型)进行视觉分析的dsh插件。A dsh plugin that allows AI to automatically invoke VLMs (multimodal models) for visual analysis.
dsh-vision-provider
作者 libinyam
Config-only DeepSeek Harness bundle for OpenAI-compatible vision models.
qwen-mm-plugins
作者 omdsh-dev
Qwen-MM-Plugins支持
dsh-tool-vision
作者 scorp1o117
Vision model for DeepSeek Harness | DeepSeek Harness 外置视觉模型插件
dsh-deepseek-vision
作者 siegfly
Vision-language gateway plugin for DeepSeek Harness - paste an image, DeepSeek sees text
ds-vision-plugin
作者 sorwcyra
Paste images into DeepSeek Harness with a four-model vision race, OCR, and an automatic text bridge.
dsh-chat-imagine
作者 corrinehu
在 DSH 聊天窗口自动调用生图工具(API 渠道,或本机 CLI:已支持mmx / codex / agy)并展示图片。
dsh-llm-vision-bridge
作者 einskyle
DeepSeek vision bridge for dsh: route image attachments to a vision model (Qwen3-VL via pi-ai/llama.cpp) and continue on a text-only LLM (DeepSeek)
free-vision-skill
作者 niyongsheng
Local‑only vision skill for macOS 本地化识图技能
dsh-ernie-image
作者 omdsh-dev
百度 ERNIE-Image-Turbo 文生图:宿主工具生成图片落盘并注册为会话附件,浏览器配置卡与生成画廊面板
