返回目录

GITHUB TOPIC

vision

57 个项目包含此标签

57 个项目

GitHub Topic 精确匹配

文件与数据技能精选

modlens

liustack

The first vision plugin for DeepSeek Harness, and the vision bridge for every text-only coding agent. Paste an image, get structured JSON evidence (OCR, layout, semantics). | 全网第一个 DeepSeek Harness 视觉插件,为 DeepSeek、GLM 等纯文本模型外挂视觉能力,粘贴图片即得结构化 JSON 证据(OCR、版面、语义)。

1.4kTypeScript2026-02-22
文件与数据插件

Eyes for text-only DeepSeek Harness agents: built-in free vision chain (no key) + pixel-level vision tools (Q&A, grounding, crop, pixel diff, colors, OCR, SVG trace, cutout, screenshots). One-command install, no Python, image turns work like ordinary tool-calling turns.

71JavaScript2026-08-14
文件与数据插件

dsh-vision

linenxi-ctrl

为 DeepSeek Harness 增加外挂识图模型:圆形鲸鱼按钮、发送图片识图自动回传、模型自主截图+识图工具、多协议自动适配、小白一键安装(未装 Node.js 自动下载)

10JavaScript2026-08-14
文件与数据插件

DeepSeek Harness 插件:DeepSeek 大脑 + 自动识图。GUI 附加图片自动经 OpenAI 兼容 VLM 转译成文字后交给 DeepSeek 作答;支持百炼/智谱/OpenRouter 等任意 OpenAI 兼容端点(默认 qwen3.7-flash),无 key 自动探测本地 Ollama(图片不出本机);安装时有一问式确认

7JavaScript2026-08-13
文件与数据插件

DSH 插件:图片与文件直达纯文本模型——图片保留原生附件体验,PDF/Office/压缩包/视频/音频显示为附件栏方块,点击发送时自动转为工作区路径,配合 dsh-vision-toolkit 粘贴即看图。A DSH plugin that delivers images AND files to text-only models as workspace paths: images keep the native attachment UI, other files show as square chips in the rail, paths append on send — pairs with dsh-vision-toolkit.

7JavaScript2026-08-14
文件与数据插件

Dsh-visual-plugin.Give your text-only model eyes: forward user images to any OpenAI-compatible vision model and see the results in a Web UI right panel

5TypeScript2026-08-14
文件与数据插件

DeepEye vision plugin for DeepSeek Harness (DSH): image description, OCR, VQA, UI layout, and clipboard analysis.

4TypeScript2026-08-14
文件与数据插件

dsh-vision

Terry12138qy

DeepSeek Harness 识图插件:为不具备原生识图能力的模型提供识图能力(阿里云百炼 qwen3.5-omni-plus,失败自动切换智谱 glm-4.6v-flash)。由 claude-vision-skill 移植适配。 | Vision tool for DeepSeek Harness

3JavaScript2026-08-14
文件与数据插件

本地 OCR 插件:让纯文本生成 LLM 也能读懂图片 | Local OCR plugin: give text-only generative LLMs the ability to read images

3TypeScript2026-08-15
文件与数据插件

给 DeepSeek 安装一双眼睛和一支画笔:会话里直接贴截图/图片,GLM 视觉模型先精确转写图片内容(报错信息、代码、界面逐字保留),然后 DeepSeek 继续处理你的问题——同一轮完成,全程无感;需要配图时,DeepSeek 自动调用文生图后端出图并显示在会话中。

3TypeScript2026-08-14
文件与数据插件

Native interactive visual-reasoning plugin for DeepSeek Harness: precise pixel grounding (SOM grid / zoom / annotate / measure / diff / color / OCR) + MiMo V2.5 multimodal backend, zero external MCP servers.

3JavaScript2026-08-14
文件与数据插件

dsh-vision

237229953-create

DSH plugin: text-only models (e.g. DeepSeek-V4) automatically see images via a vision model. Official surface-replace, cache-friendly, human transcript untouched. 纯文本模型自动识图桥

2JavaScript2026-08-14
文件与数据技能

dsh-media-skills

akqwpeter-prog

Free image reading & generation for DeepSeek Harness — paste an image into any chat, even text-only sessions. 免费读图·生图 · 9 种语言 · 无 Key 入库

2Python2026-08-14
文件与数据插件

GLM-4.6V 图像理解 MCP:识图/OCR/图表解析,原生接入 DeepSeek Harness(dsh-mcp-client),也兼容 Codex/Cline 等

2Python2026-08-14
Agent 与会话插件

Auxiliary models for DeepSeek Harness: vision understanding and context compression through dedicated model routes.

2TypeScript2026-08-14
Agent 与会话插件

DSH plugin: dispatch image-recognition tasks to an opencode-go mimo-v2.5 subagent via system-prompt injection

2JavaScript2026-08-14
文件与数据插件

DSH 插件:让纯文本模型也能看图。Web 端直接粘贴图片即可发送,无需指定图片路径;模型自主调用视觉技能查看,多模态模型原生直通,零skill绑定。

2JavaScript2026-08-14
文件与数据插件

DSH 视觉桥接插件:让无视觉能力的主模型看图(会话收图 + 自动转文字 + view_image 工具)

2JavaScript2026-08-14
文件与数据插件

Vision-augmented DeepSeek adapter plugin for DeepSeek Harness: a vision-capable model describes image input, then a text-only DeepSeek model reasons over the description

2TypeScript2026-08-14
模型与 MCP插件

dsh-qwen-mm

RRRosmontis

Qwen-MM-Plugins integration bundle for DeepSeek Harness (dsh) — multimodal MCP tools (vision, OCR, ASR, search, video, Blender, FreeCAD) + image attachment bridge. 让 DeepSeek Harness 原生支持多模态。

2TypeScript2026-08-14
文件与数据插件

A see_image vision tool plugin for DeepSeek Harness — describe images through any OpenAI-compatible vision model (GitHub Copilot, OpenAI, Ollama, vLLM, LM Studio).

2JavaScript2026-08-14
文件与数据插件

dsh-vision-bridge

Xieweikang123

Give a text-only dsh model eyes: pasted images recognized into text via an OpenAI-compatible vision endpoint.

2JavaScript2026-08-14
界面增强插件

dsh-ui-spec

yumimanji

DeepSeek Harness plugin: turn UI screenshots into structured, implementation-grade web frontend specs. Deterministic geometry (sharp) + optional vision-model semantics, merged into one JSON + Markdown spec.

2TypeScript2026-08-14
文件与数据插件

dsh-vision

sjakdhasdh

Vision tool plugin for DeepSeek Harness (DSH): give text-only models like deepseek-v4-flash image recognition via Alibaba Bailian / any OpenAI-compatible vision API. 给 DeepSeek Harness 无识图能力模型加识图工具。

1TypeScript2026-08-14
文件与数据插件

dsh-plugins

Bernardxu123

DeepSeek Harness (dsh) 插件集合: dsh-sensenova-image 生图 + dsh-vision 看图, 克隆即装

1JavaScript2026-08-14
文件与数据插件

Community plugin for DeepSeek Harness: give text-only models eyes - paste images natively, described via an OpenAI-compatible vision API

1TypeScript2026-08-14
文件与数据插件

dsh-sight

Fu3rte

Plug-in vision for text-only DeepSeek Harness (dsh) models: built-in free/cheap VLM presets + multi-image batch analysis

1JavaScript2026-08-14
文件与数据插件

dsh-img

gmleong

Give text-only models eyes: analyze_image tool for DeepSeek Harness, backed by free Chinese vision APIs (GLM-4V-Flash / Qwen-VL) or any OpenAI-compatible endpoint. 给纯文本模型装上眼睛的 dsh 插件。

1JavaScript2026-08-14
文件与数据插件

DeepSeek Harness plugin that bridges session images to pluggable vision APIs while keeping DeepSeek as the primary model.

1TypeScript2026-08-14
文件与数据插件

让纯文本模型通过桌面豆包看见聊天图片的 DeepSeek Harness 宿主插件(CDP 桥接,全预设生效,识别可取消)

1JavaScript2026-08-14
消息通讯渠道适配

DSH_plugins_4U

honghudavy-star

DSH 自建插件集合:微信桥接器 + GUI 微信入口补丁,一键安装

1JavaScript2026-08-14
文件与数据插件

DeepSeek Harness (dsh) bundle plugin: vision_analyze tool lets text-only LLM agents read images via SenseNova VLM, with Schemastery config and single-source credentials

1JavaScript2026-08-14
文件与数据插件

Paste images into DeepSeek Harness with a four-model vision race, OCR, and an automatic text bridge.

1JavaScript2026-08-14
文件与数据插件

dsh-computer-use

xiaoheizi1212

Model-agnostic Computer Use for DeepSeek Harness: isolated browser, Windows native helper, third-party vision perception, and a Chrome Cookie Bridge.

1TypeScript2026-08-14
文件与数据插件

eagleeye-mcp

baimaomaomao556

EagleEye MCP — pixel-accurate visual toolbox for Agents (screenshot, measure, OCR, regression)

0Python2026-08-15
文件与数据插件

DSH-AUX

DoloresCaritasAngelus

Auxiliary model system for DeepSeek Harness: unified aux-LLM routing (per-task model, timeout, concurrency, failure cooldown, main-model fallback) + vision_analyze / web_extract / compress_text tools, settings page, and session image lifecycle cleanup.

0JavaScript2026-08-15
文件与数据技能

让 DSH 里任何模型(包括 DeepSeek 纯文本模型)都能识别图片:识图技能 + 幂等宿主补丁,装上即用 / Let every DSH model see images: vision skill + idempotent host patch.

0JavaScript2026-08-15
文件与数据渠道适配

Image message bridge for text-only models in DeepSeek Harness (dsh): image blocks → text placeholder + local path, vision via qwen script

0JavaScript2026-08-14
文件与数据插件

dsh-vision-relay

junhongchashui

零修改、零切换的 DeepSeek Harness 视觉能力插件:纯文本模型粘贴即读图片,云端 + 本地 Ollama 双后端自动切换,ModLens v2 风格结构化证据输出。

0JavaScript2026-08-14
文件与数据插件

DeepSeek Harness 识图插件:保持 DeepSeek 对话,15+ 供应商视觉模型把图片转译为文字,可在 设置→插件 配置

0JavaScript2026-08-15
文件与数据插件

让 DeepSeek Harness 获得"看图"能力,自动识别图片真实格式,经任意 OpenAI 兼容视觉模型返回详细文字描述。

0JavaScript2026-08-14
文件与数据插件

一个可以让没有视觉的大模型拥有视觉能力的插件(当然,是通过外挂视觉模型实现的)

0JavaScript2026-08-15
模型与 MCP插件

vision_kit

Seom-ingit

Make your AI agent a math tutor. Structured extraction of vectors, matrices & geometry from math figures, with dimension-consistency + geometric self-check. Vision plugins for DeepSeek Harness, opencode (MCP) & CLI. Verify, don't believe.

0Python2026-08-14
文件与数据插件

Vision for text-only LLMs in DeepSeek Harness (DSH): describe images / OCR / VQA via free Gemini & GLM vision APIs

0JavaScript2026-08-14
文件与数据插件

DSH bundle: Qwen multimodal bridge — vision (qwen3-vl), speech-to-text (qwen3-asr), text-to-image (qwen-image), for DeepSeek Harness

0JavaScript2026-08-14