返回目录

GITHUB TOPIC

multimodal

28个项目包含此标签

28 个项目

GitHub Topic 精确匹配

安全与治理

dsh-crew

ZSeven-W

DeepSeek Harness (DSH) plugin: dispatch work to DSH agents from Claude Code / Codex — native subagent progress, in-host worker sessions with per-tier presets, and a multimodal bridge that lends the text-only harness vision and image generation.

已验证
ai-agentsclaude-codecodexcoding-agent
文件与数据

dsh-vision

oil-oil

Near-native image understanding for DeepSeek Harness

已验证
deepseek-harnessdsh-pluginimage-understandingmultimodal
文件与数据

DeepSeek Harness 插件:DeepSeek Pro 大脑 + 自动识图。GUI 附加图片默认经官方 deepseek-v4-flash-vision-exp 原生识图,转译成文字后交给 DeepSeek 作答(纯文本的 V4-Pro 也能看图);支持百炼/智谱/OpenRouter 等任意 OpenAI 兼容 VLM,无 key 自动探测本地 Ollama;安装时有一问式确认

已验证
dashscopedeepseek-harnessdsh-pluginimage-understanding
开发工具

DeepSeek Harness 第三方 API 与自定义模型设置插件:支持请求头、User-Agent、模型列表、图像输入和思考等级 | WebUI plugin for third-party APIs and custom models with request headers, image input, and reasoning levels

已验证
deepseek-harnessdshdsh-pluginmodel-provider
文件与数据

DeepEye vision plugin for DeepSeek Harness (DSH): image description, OCR, VQA, UI layout, and clipboard analysis.

已验证
cordisdeepseek-harnessdshdsh-plugin
文件与数据

Vision routing and image generation for DeepSeek Harness through a fixed Mix model.

已验证
deepseek-harnessdsh-plugingpt-image-2image-generation
文件与数据

给 DeepSeek 安装一双眼睛和一支画笔:会话里直接贴截图/图片,GLM 视觉模型先精确转写图片内容(报错信息、代码、界面逐字保留),然后 DeepSeek 继续处理你的问题——同一轮完成,全程无感;需要配图时,DeepSeek 自动调用文生图后端出图并显示在会话中。

已验证
deepseek-harnessdshdsh-plugindsh-plugins
开发工具

visual-review

wang-bool

DeepSeek Harness plugin for image upload, in-chat rendering, and vision analysis with cloud or local models. 图像上传、显示与解析。

已验证
deepseek-harnessdshdsh-pluginimage-analysis
开发工具

DeepSeek Herness plugin. dsh插件,支持在创建自定义模型时手动选择模型能力,比如模型是否支持图片输入等。DeepSeek Herness plugin. dsh插件,支持在创建自定义模型时手动选择模型能力,比如模型是否支持图片输入等。DeepSeek Herness plugin. dsh插件,支持在创建自定义模型时手动选择模型能力,比如模型是否支持图片输入等。The DSH plugin allows users to manually configure model capabilities when creating a custom model, such as image input support.

已验证
deepseek-harnessdsh-pluginmodel-capabilitiesmodel-configuration
Agent 与会话

Model catalog, portraits, Agent selection, and multimodal runtimes for DeepSeek Harness

已验证
aideepseek-harnessdshdsh-plugin
文件与数据

dsh-xiapan-media

dongsheng123132

Native vision, gpt-image-2 and Seedance plugins for DeepSeek Harness via Xiapan Cloud

已验证
deepseek-harnessdsh-pluginimage-generationmultimodal
文件与数据

dsh-vision-bridge

GooDAnDReaDY

Universal vision bridge for DeepSeek Harness: attachments with native models, 40+ tools, PDF/OCR/diagrams.

已验证
deepseek-harnessdshdsh-pluginmultimodal
文件与数据

dsh-open-eyes

hyper-dsh-plugins

A lightweight DeepSeek Harness vision delegation tool for text-only routes, with native OpenAI Responses, Chat Completions, and Anthropic Messages adapters.

已验证
anthropicdeepseek-harnessdsh-pluginmultimodal
文件与数据

DeepSeek Harness vision plugin: analyze_image (structured OCR evidence) + capture_image (USB camera visual loop). 摄像头视觉闭环 + 结构化证据,支持 Ollama / DeepSeek / Xiaomi 三后端。

已验证
cameradeepseek-harnessdsh-pluginimage-to-text
文件与数据

dsh-vision

reimu-create

DSH plugin: text-only models (e.g. DeepSeek-V4) automatically see images via a vision model. Official surface-replace, cache-friendly, human transcript untouched. 纯文本模型自动识图桥

已验证
deepseek-harnessdsh-pluginmultimodalvision
文件与数据

Silent vision bridge for DeepSeek Harness: route chat images to a fixed vision model, preserve UI originals, and reuse observations across compaction and restarts.

已验证
deepseek-harnessdshdsh-pluginimage
文件与数据

dsh-eye-vision

AlloyPlane

该仓库暂未提供项目说明。

已验证
aideepseek-harnessdshdsh-plugin
文件与数据

统一接入 Image2 与 Nano Banana 四款模型,覆盖文生图、多参考图编辑、2K/4K 输出、顺序批量任务、默认模型持久化和脱敏 Key 配置。

已验证
88apideepseek-harnessdshdsh-plugin
文件与数据

dsh-mindseye

kanchengw

Plug-in vision for text-only models on DSH, with native interaction for image understanding and generation, and GUI automation, through layered evidence memory and cache.

已验证
agentdeepseek-harnessdsh-plugingui-automation
文件与数据

dsh-eyes

Leeminjing

Give text-only DeepSeek models on-demand vision: upload images, DeepSeek answers by calling a view_image tool backed by any OpenAI-compatible vision endpoint (Qwen/DashScope by default).

已验证
dashscopedeepseek-harnessdsh-pluginmultimodal
文件与数据

Transparent image preprocessing route for DeepSeek Harness

已验证
ai-agentscordisdeepseekdeepseek-harness
文件与数据

DSH plugin: text-only chat models get local Ollama VL image descriptions (qwen3-vl:8b), VRAM cooling

已验证
deepseek-harnessdshdsh-pluginmultimodal
文件与数据

dsh-qwen-mm

RRRosmontis

Qwen-MM-Plugins integration bundle for DeepSeek Harness (dsh) — multimodal MCP tools (vision, OCR, ASR, search, video, Blender, FreeCAD) + image attachment bridge. 让 DeepSeek Harness 原生支持多模态。

已验证
agentaideepseekdeepseek-harness
文件与数据

Vision for DeepSeek Harness agents — paste images in the Web composer, delegate reads to Kimi/MiniMax vision routes on isolated contexts; zero image bytes in the main session

已验证
ai-agentcordisdeepseekdeepseek-harness
文件与数据

dsh-vision-link

sprainJinyu

Route-preserving image understanding for text-only models in DeepSeek Harness (DSH).

已验证
deepseek-harnessdsh-pluginimage-understandingjavascript
开发工具

Bring your Grok subscription into DSH as an ACP subagent, extending native images with audio and video tools.

已验证
acpaudiodeepseek-harnessdsh-plugin
模型与 MCP

TokenLab native-protocol model provider, multimodal tools, and async tasks for DeepSeek Harness

已验证
async-aideepseek-harnessdsh-pluginmcp
文件与数据

um-dsh-azimg

UnforgetMemory

Give DeepSeek Harness (DSH) agents eyes: local image analysis via vision models — um_analyze_img tool, model-capability recognition, hot-switch settings UI.

已验证
ai-agentscordisdeepseek-harnessdsh