返回目录

GITHUB TOPIC

vision

46个项目包含此标签

46 个项目

GitHub Topic 精确匹配

文件与数据

dsh-vision

oil-oil

Near-native image understanding for DeepSeek Harness

已验证
deepseek-harnessdsh-pluginimage-understandingmultimodal
文件与数据

Free image reading & generation for DeepSeek Harness (rc.7 / rc.8 / v0.1.1-rc.1 / rc.2 / v0.1.2-alpha.3) — paste-image reading with auto vision transcription, DeepSeek-V4-Flash-Vision-Exp / GLM-4V-Flash / SenseNova / Gemini failover, Kolors + U1 Fast generation. No keys in repo.

已验证
agent-skillsdeepseek-harnessdsh-pluginimage-generation
文件与数据

DeepSeek Harness 插件:DeepSeek Pro 大脑 + 自动识图。GUI 附加图片默认经官方 deepseek-v4-flash-vision-exp 原生识图,转译成文字后交给 DeepSeek 作答(纯文本的 V4-Pro 也能看图);支持百炼/智谱/OpenRouter 等任意 OpenAI 兼容 VLM,无 key 自动探测本地 Ollama;安装时有一问式确认

已验证
dashscopedeepseek-harnessdsh-pluginimage-understanding
文件与数据

DSH 插件:图片与文件直达纯文本模型——图片保留原生附件体验,PDF/Office/压缩包/视频/音频显示为附件栏方块,点击发送时自动转为工作区路径,配合 dsh-vision-toolkit 粘贴即看图。A DSH plugin that delivers images AND files to text-only models as workspace paths: images keep the native attachment UI, other files show as square chips in the rail, paths append on send — pairs with dsh-vision-toolkit.

已验证
agentattachmentsdeepseekdeepseek-harness
文件与数据

一个工具 = MiniMax 全部多模态能力:DSH 纯文本模型看图/画图/生视频/说话/唱歌/翻唱/搜索/查额度 | One mmx_bridge tool = all MiniMax multimodal (VLM/image/video/speech/music/cover/search/quota) for DeepSeek Harness (DSH)

已验证
agent-toolai-agentcordisdeepseek-harness
文件与数据

专供 deepseek-v4-flash-vision-exp 的高清识图增强插件:放宽 DSH 图片限制 + highres_read 分块识图工具。

已验证
deepseek-harnessdsh-pluginimage-recognitionvision
文件与数据

为DSH(DeepSeek Harness)量身打造的视觉插件,现已支持agent调用图片显示/Vision plugin for DSH(DeepSeek Harness),support Proactive Image Display.

已验证
deepseek-harnessdshdsh-pluginfree
开发工具

DSH 轻量截图插件:轻 : 纯 PowerShell 实现,零依赖、零二进制;截图能力独立维护,不随任何上游更新而失效。 摆 :一键全屏即拍;框选前整个桌面保持可操作,所有窗口像布置画面一样自由移动、缩放;鼠标悬停任意窗口即亮起吸附边框,点一下直接截该窗口,被遮挡也能拿到完整内容,所得即所见。 自助:Agent 可随时自行截屏;接上 modlens(选装),截屏 + 识图一步产出结构化内容(OCR/版面/语义),纯文本模型也能消费。

已验证
deepseek-harnessdsh-pluginscreenshotvision
开发工具

DeepSeek Harness plugin that bridges session images to pluggable vision APIs while keeping DeepSeek as the primary model.

已验证
deepseekdeepseek-aideepseek-harnessdeepseek-harness-desktop
文件与数据

Codex-style attachment formats for the DeepSeek Harness Web GUI: PDF text-layer extraction, Office text extraction, scanned-PDF OCR, long-document spill + index cards, image-to-PNG.

已验证
attachmentcordisdeepseekdeepseek-harness
文件与数据

Local‑only vision skill for macOS 本地化识图技能dsh-plugin

已验证
deepseek-harnessdsh-plugindsh-skillocr
文件与数据

Vision for text-only LLMs in DeepSeek Harness (DSH): describe images / OCR / VQA via free Gemini & GLM vision APIs

已验证
deepseek-harnessdshdsh-plugingemini
文件与数据

DeepSeek Harness usage dashboard with API balance, daily spend, external vision-call accounting, per-model stats, call logs, cache rate, TTFT, and CSV export.

已验证
dashboarddeepseek-harnessdshdsh-plugin
文件与数据

DeepEye vision plugin for DeepSeek Harness (DSH): image description, OCR, VQA, UI layout, and clipboard analysis.

已验证
cordisdeepseek-harnessdshdsh-plugin
文件与数据

Vision routing and image generation for DeepSeek Harness through a fixed Mix model.

已验证
deepseek-harnessdsh-plugingpt-image-2image-generation
文件与数据

给 DeepSeek 安装一双眼睛和一支画笔:会话里直接贴截图/图片,GLM 视觉模型先精确转写图片内容(报错信息、代码、界面逐字保留),然后 DeepSeek 继续处理你的问题——同一轮完成,全程无感;需要配图时,DeepSeek 自动调用文生图后端出图并显示在会话中。

已验证
deepseek-harnessdshdsh-plugindsh-plugins
开发工具

visual-review

wang-bool

DeepSeek Harness plugin for image upload, in-chat rendering, and vision analysis with cloud or local models. 图像上传、显示与解析。

已验证
deepseek-harnessdshdsh-pluginimage-analysis
文件与数据

图片识别插件 for DeepSeek Harness:自动判断当前模型识图能力,支持多供应商视觉模型管理与检测

已验证
computer-visioncordis-plugindeepseek-harnessdsh-plugin
文件与数据

dsh-vision-skill

DDDFXYqiming

Vision skill plugin for DeepSeek Harness (image analysis and OCR)

已验证
agent-skillsdeepseek-harnessdshdsh-plugin
文件与数据

dsh-xiapan-media

dongsheng123132

Native vision, gpt-image-2 and Seedance plugins for DeepSeek Harness via Xiapan Cloud

已验证
deepseek-harnessdsh-pluginimage-generationmultimodal
文件与数据

dsh-tool-eyes

go-farther-and-farther

DeepSeek Harness (DSH) 本地视觉眼睛插件:screen 工具(截图/图片交给本地视觉模型描述)+ ocr 工具(Windows 内置 OCR 逐字提取文字)。零云端、OCR 零 GPU、图片不出本机。

已验证
deepseek-harnessdsh-pluginocrvision
文件与数据

dsh-vision-bridge

GooDAnDReaDY

Universal vision bridge for DeepSeek Harness: attachments with native models, 40+ tools, PDF/OCR/diagrams.

已验证
deepseek-harnessdshdsh-pluginmultimodal
文件与数据

dsh-open-eyes

hyper-dsh-plugins

A lightweight DeepSeek Harness vision delegation tool for text-only routes, with native OpenAI Responses, Chat Completions, and Anthropic Messages adapters.

已验证
anthropicdeepseek-harnessdsh-pluginmultimodal
文件与数据

DeepSeek Harness 识图插件:保持 DeepSeek 对话,15+ 供应商视觉模型把图片转译为文字,可在 设置→插件 配置

已验证
deepseek-harnessdsh-pluginvision
文件与数据

DeepSeek Harness vision plugin: analyze_image (structured OCR evidence) + capture_image (USB camera visual loop). 摄像头视觉闭环 + 结构化证据,支持 Ollama / DeepSeek / Xiaomi 三后端。

已验证
cameradeepseek-harnessdsh-pluginimage-to-text
文件与数据

DSH 投屏控制的共用核心:界面、设备列表、投屏面板与 AI 工具(配合其他 provider 使用)

已验证
deepseek-harnessdshdsh-pluginscrcpy
文件与数据

dsh-vision

reimu-create

DSH plugin: text-only models (e.g. DeepSeek-V4) automatically see images via a vision model. Official surface-replace, cache-friendly, human transcript untouched. 纯文本模型自动识图桥

已验证
deepseek-harnessdsh-pluginmultimodalvision
文件与数据

dsh-vision

sjakdhasdh

Vision tool plugin for DeepSeek Harness (DSH): give text-only models like deepseek-v4-flash image recognition via Alibaba Bailian / any OpenAI-compatible vision API. 给 DeepSeek Harness 无识图能力模型加识图工具。

已验证
ai-agentdeepseekdeepseek-harnessdsh-plugin
文件与数据

Silent vision bridge for DeepSeek Harness: route chat images to a fixed vision model, preserve UI originals, and reuse observations across compaction and restarts.

已验证
deepseek-harnessdshdsh-pluginimage
文件与数据

dsh-eye-vision

AlloyPlane

该仓库暂未提供项目说明。

已验证
aideepseek-harnessdshdsh-plugin
开发工具

DSH plugin: auto-detect and configure model capabilities (reasoningEfforts + input modalities) for llm-pi-ai. Successor to dsh-reasoning-efforts.

已验证
deepseek-harnessdshdsh-plugininput-modalities
文件与数据

aura-vision

Ck-epsilon

Aura Vision - free vision OCR plugin for DeepSeek Harness web profile (permanent bundle, GLM-4V-Flash free tier, tile-based long-document recognition)

已验证
deepseek-harnessdsh-pluginglm-4vocr
文件与数据

dsh-image-mmx

fengs2021

给 DSH 文本模型装眼睛:图片自动调用 mmx(MiniMax VLM)识别,识别结果注入模型上下文

已验证
deepseek-harnessdsh-pluginvision
文件与数据

dsh-image-unlock

FrostLeafKEE

DeepSeek Harness 插件:解除 Web GUI 图片输入限制,图片附件文本化后交给 vision skill 识图 | dsh plugin that lifts the image-input gate and hands attachments to a vision skill

已验证
deepseek-harnessdsh-pluginpluginskill
文件与数据

DeepSeek Harness 图像理解插件 · 8 种分析模式 · 支持任意API接口 · 内置免费视觉模型 | DeepSeek Harness vision plugin · 8 analysis modes · works with any OpenAI- or Anthropic-compatible API · built-in free vision model

已验证
deepseekdeepseek-harnessdsh-pluginfree
文件与数据

DeepSeek Harness(DSH)视觉插件:Edge+豆包网页版识图,零成本免 API Key。通用识图 + 数学建模图专项(几何/流程图/图表/表格/公式)+ 不确定项澄清闭环。Vision plugin for DeepSeek Harness: image understanding via Edge + Doubao Web, zero cost, no API key. General recognition + math-modeling diagrams (geometry/flowcharts/charts/tables/formulas) + clarify loop for uncertainties.

已验证
deepseek-harnessdsh-pluginvision
文件与数据

dsh-vision

kaaaaahn

DSH 本地视觉能力插件:macOS Vision OCR + ollama qwen3-vl 语义描述 + 上传图片桥接

已验证
deepseek-harnessdsh-pluginlocal-aiocr
文件与数据

dsh-mindseye

kanchengw

Plug-in vision for text-only models on DSH, with native interaction for image understanding and generation, and GUI automation, through layered evidence memory and cache.

已验证
agentdeepseek-harnessdsh-plugingui-automation
文件与数据

dsh-pro-vision

lasdrder0705

DSH plugin: let DeepSeek-V4-Pro use V4-Flash-Vision-Exp for attached images. Install: dsh plugin --profile web add github:lasdrder0705/dsh-pro-vision

已验证
deepseek-harnessdsh-pluginvision
文件与数据

dsh-eyes

Leeminjing

Give text-only DeepSeek models on-demand vision: upload images, DeepSeek answers by calling a view_image tool backed by any OpenAI-compatible vision endpoint (Qwen/DashScope by default).

已验证
dashscopedeepseek-harnessdsh-pluginmultimodal
文件与数据

DSH plugin: text-only chat models get local Ollama VL image descriptions (qwen3-vl:8b), VRAM cooling

已验证
deepseek-harnessdshdsh-pluginmultimodal
文件与数据

Vision for DeepSeek Harness agents — paste images in the Web composer, delegate reads to Kimi/MiniMax vision routes on isolated contexts; zero image bytes in the main session

已验证
ai-agentcordisdeepseekdeepseek-harness
文件与数据

dsh-vision-link

sprainJinyu

Route-preserving image understanding for text-only models in DeepSeek Harness (DSH).

已验证
deepseek-harnessdsh-pluginimage-understandingjavascript
文件与数据

该仓库暂未提供项目说明。

已验证
deepseek-harnessdeepseek-harness-plugindshdsh-plugin
文件与数据

DeepSeek Harness 视觉补全:孪生路由解锁原生图片体验,本地 Ollama 请求层看图,零云端依赖。Vision twin + local agentic vision tools for DeepSeek Harness.

已验证
deepseek-harnessdsh-pluginollamavision
文件与数据

DSH Computer Use:让 Agent 看见并操作 Windows 桌面(Windows 原生 / WSL 自动识别);视觉通道 + 鼠标键盘 + Codex 风格蓝色覆盖层,Esc 随时中止。

已验证
automationcomputer-usedeepseek-harnessdsh