VultrVultr
返回目录

GITHUB TOPIC

multimodal

42个项目包含此标签

42 个项目

GitHub Topic 精确匹配

文件与数据插件

Eyes for text-only DeepSeek Harness agents: built-in free vision chain (no key) + pixel-level vision tools (Q&A, grounding, crop, pixel diff, colors, OCR, SVG trace, cutout, screenshots). One-command install, no Python, image turns work like ordinary tool-calling turns.

待结构检查
deepseek-harnessdshdsh-pluginmultimodal
文件与数据插件

dsh-vision

oil-oil

Near-native image understanding for DeepSeek Harness

待结构检查
deepseek-harnessdsh-pluginimage-understandingmultimodal
模型与 MCP插件

dsh-AuthInOne

Stormycry-cryp

Self-contained DeepSeek Harness (DSH) plugin for Provider/Auth login, model switching, image fallback, token/cost analytics, and same-port Web restart. Useful? A star helps.

待结构检查
cost-attributioncost-trackingcustom-apideepseek-harness
文件与数据插件

DeepSeek Harness 插件:DeepSeek 大脑 + 自动识图。GUI 附加图片自动经 OpenAI 兼容 VLM 转译成文字后交给 DeepSeek 作答;支持百炼/智谱/OpenRouter 等任意 OpenAI 兼容端点(默认 qwen3.7-flash),无 key 自动探测本地 Ollama(图片不出本机);安装时有一问式确认

待结构检查
dashscopedeepseek-harnessdsh-pluginimage-understanding
文件与数据插件

Dsh-visual-plugin.Give your text-only model eyes: forward user images to any OpenAI-compatible vision model and see the results in a Web UI right panel

待结构检查
deepseekdeepseek-harnessdsh-pluginimage-description
文件与数据完整应用

Vision-language gateway plugin for DeepSeek Harness - paste an image, DeepSeek sees text

非插件验证范围
coding-agentdeepseekdeepseek-harnessdsh
文件与数据插件

DeepSeek Harness plugin: describe_image — give a text-only model vision through an OpenAI-compatible VLM endpoint

待结构检查
deepseekdeepseek-harnessdescribe-imagedsh
文件与数据插件

给 DeepSeek 安装一双眼睛和一支画笔:会话里直接贴截图/图片,GLM 视觉模型先精确转写图片内容(报错信息、代码、界面逐字保留),然后 DeepSeek 继续处理你的问题——同一轮完成,全程无感;需要配图时,DeepSeek 自动调用文生图后端出图并显示在会话中。

待结构检查
deepseek-harnessdshdsh-plugindsh-plugins
文件与数据插件

DeepEye vision plugin for DeepSeek Harness (DSH): image description, OCR, VQA, UI layout, and clipboard analysis.

待结构检查
cordisdeepseek-harnessdshdsh-plugin
文件与数据插件

该仓库暂未提供项目说明。

待结构检查
deepseekdeepseek-harnessdshdsh-plugin
文件与数据插件

Config-only DeepSeek Harness bundle for OpenAI-compatible vision models.

待结构检查
deepseek-harnessdshdsh-pluginmultimodal
文件与数据插件

Paste images into DeepSeek Harness with a four-model vision race, OCR, and an automatic text bridge.

待结构检查
deepseekdeepseek-harnessdsh-pluginimage-to-text
文件与数据插件

dsh-xiapan-media

dongsheng123132

Native vision, gpt-image-2 and Seedance plugins for DeepSeek Harness via Xiapan Cloud

待结构检查
deepseek-harnessdsh-pluginimage-generationmultimodal
文件与数据插件

dsh-vision

reimu-create

DSH plugin: text-only models (e.g. DeepSeek-V4) automatically see images via a vision model. Official surface-replace, cache-friendly, human transcript untouched. 纯文本模型自动识图桥

待结构检查
deepseek-harnessdsh-pluginmultimodalvision
文件与数据插件

MiniMax multimodal capability hub for DeepSeek Harness (DSH): image understanding (VLM), text/image-to-video, speech, music, audio cover, web search, quota — one mmx_multimodal model tool wrapping the mmx-cli.

待结构检查
agent-toolai-agentcordisdeepseek-harness
文件与数据插件

Vision routing and image generation for DeepSeek Harness through a fixed Mix model.

待结构检查
deepseek-harnessdsh-plugingpt-image-2image-generation
模型与 MCP插件

A lightweight DeepSeek Harness vision delegation tool for text-only routes, with native OpenAI Responses, Chat Completions, and Anthropic Messages adapters.

待结构检查
anthropicdeepseek-harnessdsh-pluginmultimodal
文件与数据插件

Transparent image preprocessing route for DeepSeek Harness

待结构检查
ai-agentscordisdeepseekdeepseek-harness
模型与 MCP插件

dsh-qwen-mm

RRRosmontis

Qwen-MM-Plugins integration bundle for DeepSeek Harness (dsh) — multimodal MCP tools (vision, OCR, ASR, search, video, Blender, FreeCAD) + image attachment bridge. 让 DeepSeek Harness 原生支持多模态。

待结构检查
agentaideepseekdeepseek-harness
模型与 MCP插件

DeepSeek Harness 的 Qwen-MM-Plugins 集成插件:12 个多模态 MCP 工具(视觉/OCR/定位/ASR/音视频)、Web 设置页(粘贴 Qwen API Key 即用)、内置技能与一键安装器

待结构检查
asrdeepseek-harnessdsh-pluginmcp
模型与 MCP插件

A see_image vision tool plugin for DeepSeek Harness — describe images through any OpenAI-compatible vision model (GitHub Copilot, OpenAI, Ollama, vLLM, LM Studio).

待结构检查
deepseek-harnessdshdsh-plugingithub-copilot
生活娱乐插件

dsh-voice

zhuiyueya

Voice for DeepSeek Harness(dsh) — speech-to-text input + read-aloud TTS for text-only DeepSeek, zero API key.

待结构检查
ai-agentsdeepseekdeepseek-harnessdeepseek-harness-plugin
开发工具插件

deepsee

chang416

Vision + smart model routing for DeepSeek Harness. Gemini sees. DeepSeek codes.

待结构检查
ai-agentsai-coding-agentclaude-codecodex
文件与数据插件

DeepSeek Harness 视觉增强插件:将图片交给外部视觉模型分析,输出带坐标化视觉原语的纯文本证据,使不支持多模态的文本模型也能在对话中理解图片、截图与文档。

待结构检查
deepseek-harnessdeepseek-harness-plugindshdsh-plugin
文件与数据插件

dsh-mindseye

kanchengw

Plug-in vision for text-only models in DSH, with native multimodal interaction, layered memory cache, and intent router.

待结构检查
agentcachedeepseek-harnessdsh-plugin
文件与数据插件

dsh-eyes

Leeminjing

Give text-only DeepSeek models on-demand vision: upload images, DeepSeek answers by calling a view_image tool backed by any OpenAI-compatible vision endpoint (Qwen/DashScope by default).

待结构检查
dashscopedeepseek-harnessdsh-pluginmultimodal
文件与数据插件

给 DeepSeek Harness 纯文本模型加视觉:聊天框拖图自动分流给用户配置的视觉模型转文字,v4 不切换模型即可识图。 Vision for text-only DSH models: routes chat-box images to a user-configured vision model and returns text descriptions, deepseek-v4-pro reads images without switching.

待结构检查
codex-styledeepseek-harnessdsh-pluginimage
文件与数据插件

dsh-vision-bridge

TwistedRiCen

DSH-native Vision Evidence bridge for text-only reasoning models with native image attachments and strict multi-image validation.

待结构检查
deepseek-harnessdsh-pluginllmmultimodal
文件与数据插件

mimo-vision

wulusai2333

DeepSeek Harness (DSH) native plugin — describe_image tool: a vision bridge (image → mimo-v2.5 → text description) over the ctx.fs / ctx.credentials seams

待结构检查
agentcordisdeepseek-harnessdsh
模型与 MCP插件

Eyes for text-only DeepSeek: view_image tool (local Ollama or any OpenAI-compatible VLM) + chat image-attachment bridge — paste/drop images in the chat and the model can see them.

待结构检查
deepseek-harnessdshdsh-pluginmultimodal
文件与数据插件

Vision toolkit for DeepSeek Harness -- give text-only agents eyes

待结构检查
deepseek-harnessdeepseek-vldsh-pluginmultimodal
文件与数据技能

DeepSeek Harness (DSH) plugins. qwen-image gives a text-only coding model eyes: an image goes to a Qwen-VL route through ctx.llm and comes back as text, so DeepSeek keeps coding while Qwen looks. Pure ESM, no build permission at install. | DSH 插件集:qwen-image 让纯文本模型借千问 VL 读图,返回文本;纯 ESM,安装无需构建授权。

待结构检查
agent-skillsclaude-codecodexcoding-agent
文件与数据插件

给纯文本 DeepSeek Harness 模型加上识图能力:analyze_image 把图片转发到任意 OpenAI 兼容视觉端点 | Vision bridge for text-only DSH models

待结构检查
deepseek-harnessdsh-pluginmultimodalopenai-compatible
模型与 MCP插件

A zero-config, multi-provider vision tool for DeepSeek Harness with automatic local model discovery and privacy-aware remote fallback.

待结构检查
deepseek-harnessdsh-pluginmultimodalollama
文件与数据插件

DeepSeek Harness browser-extension edition: side panel with page-context awareness, vision-bridge multimodal image reading, and voice input

待结构检查
ai-agentbrowser-extensionchrome-extensiondeepseek
文件与数据插件

dsh-vision-relay

junhongchashui

零修改、零切换的 DeepSeek Harness 视觉能力插件:纯文本模型粘贴即读图片,云端 + 本地 Ollama 双后端自动切换,ModLens v2 风格结构化证据输出。

待结构检查
cordisdeepseek-harnessdshdsh-plugin
文件与数据插件

Vision for DeepSeek Harness agents — paste images in the Web composer, delegate reads to Kimi/MiniMax vision routes on isolated contexts; zero image bytes in the main session

待结构检查
ai-agentcordisdeepseekdeepseek-harness
文件与数据插件

Zero-core-change vision capability for DeepSeek Harness: the describe_image tool + profile bundle, installable via 'dsh plugin add'

待结构检查
ai-agentsdeepseek-harnessdshdsh-plugin
界面增强插件

DeepSeek Herness plugin. dsh插件,支持在创建自定义模型时手动选择模型能力,比如模型是否支持图片输入等。

待结构检查
deepseek-harnessdsh-pluginmodel-capabilitiesmodel-configuration
文件与数据插件

dsh-agnes-omni

wumu1111111

Agnes omni-modal plugin for DeepSeek Harness: agnes_vision (image understanding) + agnes_image (text-to-image / image-to-image) + a vision bridge that lets you send images in chat. API key via DSH credentials, never in code.

待结构检查
agnesdeepseek-harnessdshdsh-plugin
文件与数据插件

DSH bundle: Qwen multimodal bridge — vision (qwen3-vl), speech-to-text (qwen3-asr), text-to-image (qwen-image), for DeepSeek Harness

待结构检查
asrdashscopedeepseek-harnessdsh-plugin
文件与数据插件

dsh-visionary

zhuiyueya

Give text-only DeepSeek models eyes — a DeepSeek Harness plugin that transparently converts chat images into OCR text + vision-model descriptions before they reach the LLM. Configure vision backends (GLM-4V, Qwen-VL, Gemini, Ollama…) right in the Models settings page; multi-backend fallback chain, double-layer caching, no config files.

待结构检查
deepseekdeepseek-harnessdsh-pluginimage-understanding