DSH Market

47 #ocr DeepSeek Harness (DSH) Plugins

Every indexed repository carrying the GitHub topic ocr, sorted by stars. The tag is the author's own word for it — the install verdict on each card is ours.

Repositories tagged #ocr, most-starred first

M

Modlens

liustack

The first vision plugin for DeepSeek Harness, and the vision bridge for every text-only coding agent. Paste an image, get structured JSON evidence (OCR, layout, semantics). | 全网第一个 DeepSeek Harness 视觉插件,为 DeepSeek、GLM 等纯文本模型外挂视觉能力,粘贴图片即得结构化 JSON 证据(OCR、版面、语义)。

InstallableBundle1.6k
D

Dsh Vision Toolkit

Anionex

让纯文本模型更好地做视觉任务的DeepSeek Harness插件:带意图的图片问答、长截图 OCR、UI 还原等|DeepSeek Harness-native integration for agent-vision-toolkit: image Q&A, long-screenshot OCR, UI restoration, grounding, pixel diff, Artifacts, and Web UI.

Installableinstall scriptsBundle400
D

Dsh Vision Complete

Yts1919

给 DeepSeek 补上「眼睛和耳朵」的多模态视觉插件:看图 / OCR / 物体检测 / 视频理解 / 语音转写 / 截图直读,一键安装(DSH 插件)。

Not checked17
D

Dsh Vision

linenxi-ctrl

为 DeepSeek Harness 增加外挂识图模型:圆形鲸鱼按钮、发送图片识图自动回传、模型自主截图+识图工具、多协议自动适配、小白一键安装(未装 Node.js 自动下载)

InstallableBundle10
D

Dsh Vision Proxy

Flyvhidbwo

DeepSeek Harness 插件:DeepSeek 大脑 + 自动识图。GUI 附加图片自动经 OpenAI 兼容 VLM 转译成文字后交给 DeepSeek 作答;支持百炼/智谱/OpenRouter 等任意 OpenAI 兼容端点(默认 qwen3.7-flash),无 key 自动探测本地 Ollama(图片不出本机);安装时有一问式确认

Not checked8
D

Dsh Docs

Sqhao-O

Fully local document intelligence for DeepSeek Harness. Parse PDF, Office files, images, and scanned documents with offline OCR. | DeepSeek Harness 全本地文档智能插件,支持 PDF、Office、图片与离线 OCR

Not checked8
D

Dsh Windows Ocr

maxwell-feng

No description provided.

Not checked5
D

Dsh Image Bridge

kbpoyo

DSH 插件:让纯文本模型也能看图。Web 端直接粘贴图片即可发送,无需指定图片路径;模型自主调用视觉技能查看,多模态模型原生直通,零skill绑定。

Not checked4
D

Dsh Plugin Deepeye

Favio8

DeepEye vision plugin for DeepSeek Harness (DSH): image description, OCR, VQA, UI layout, and clipboard analysis.

Not checked4
C

Codex Eyes Hands

651002

专为 DeepSeek Harness 打造:把本机 Codex CLI 变成纯文本 AI agent 的眼睛和手——看图/读文件/画图/监督执行/双通道容灾

Not checked4
D

Dsh Tool Ocr

ferstar

本地 OCR 插件:让纯文本生成 LLM 也能读懂图片 | Local OCR plugin: give text-only generative LLMs the ability to read images

Not checked3
D

Dsh Deepseek Vision

Argonaut790

Image understanding, OCR, and persistent visual evidence for text-only DeepSeek Harness models

Not checked3
D

Dsh Vision Primitives

zouyuanqing

Native interactive visual-reasoning plugin for DeepSeek Harness: precise pixel grounding (SOM grid / zoom / annotate / measure / diff / color / OCR) + MiMo V2.5 multimodal backend, zero external MCP servers.

Not checked3
D

Dsh Attachment Formats

linkingoscar

Codex-style attachment formats for the DeepSeek Harness Web GUI: PDF text-layer extraction, Office text extraction, scanned-PDF OCR, long-document spill + index cards, image-to-PNG.

Not checked3
G

GrassVison

moduqishi

给纯文本大模型装上原生视觉:流式真实思考链 · 跨轮次无感重看 · 像素级证据与 SVG 图元 · OpenAI/Anthropic/Responses 三协议兼容 | Native vision for text-only LLMs: streaming real thinking chain, cross-turn re-view, pixel-level evidence & SVG primitives, OpenAI/Anthropic/Responses compatible.

Not checked3
D

DeepSeek Prism

YOGEMOW

为纯文本模型按需识图:DSH 零补丁 Cordis 插件(prism_see 工具 + 图片 VEP 降级 + 技能运行时注册)+ Codex Skill;多 Provider 视觉 API,VEP/1 低 Token 视觉证据包

Not checked3
D

Deepseek Omnimodal

good-boy4069

Open-source multimodal MCP plugin for text-only AI agents: recognize and generate images, video, and audio through Qwen/DashScope. Supports Codex, Claude Code, and DeepSeek Harness ecosystem.

Not checked3
D

DSH Plugins 4U

honghudavy-star

DSH 自建插件集合:微信桥接器 + GUI 微信入口补丁,一键安装

Not checked2
S

Shadow Vision

WardLu

Open-source MCP vision server that gives text-only LLMs and AI agents image understanding, OCR, visual analysis, UI inspection, and multimodal capabilities.

Not checked2
F

Free Vision Skill

niyongsheng

Local‑only vision skill for macOS 本地化识图技能

Not checked2
L

Locallens

uknowmyface

Local OCR for DeepSeek Harness — read text from screenshots on your Mac with Apple's Vision framework. No API key, no upload.

Not checked2
D

Ds Vision Plugin

Sorwcyra

Paste images into DeepSeek Harness with a four-model vision race, OCR, and an automatic text bridge.

Not checked2
D

Dsh PaddleOCR Skills

Aidenwu0209

PaddleOCR skills for DeepSeek Harness with native tools and GUI configuration

Not checked2
G

Glm4v Vision Mcp

ethanweave

GLM-4.6V 图像理解 MCP:识图/OCR/图表解析,原生接入 DeepSeek Harness(dsh-mcp-client),也兼容 Codex/Cline 等

Not checked2
D

Dsh Vision Bridge

Xieweikang123

Give a text-only dsh model eyes: pasted images recognized into text via an OpenAI-compatible vision endpoint.

Not checked2
D

Deepseek Harness

tylerbuilds

Local-first, safety-gated CLI and MCP harness for bounded DeepSeek batch and corpus workloads.

Not checked2
D

Dsh Mac Vision

Kevoyuan

On-device macOS OCR and Apple Vision for DeepSeek Harness — one native plugin with a bundled Skill.

Not checked1
E

Easy Vision

Koreyer

A DeepSeek Harness tool plugin that lets text-only agents "see" local images — auto-detects the real format and returns a detailed text description via any OpenAI-compatible vision model.

Not checked1
D

Doubao Vision Dsh

hawkongz

让纯文本模型通过桌面豆包看见聊天图片的 DeepSeek Harness 宿主插件(CDP 桥接,全预设生效,识别可取消)

Not checked1
D

Dsh Legal Dashboard

ByronLeeeee

Matter-aware legal workspace dashboard and document agent tools for DeepSeek Harness

Not checked1
S

Sidesight

ZhuXinAI

CLI-first vision sidecar for text-only coding agents. Analyze screenshots, diagrams, charts, UI diffs, and videos with OpenAI-compatible multimodal models.

Not checked1
D

Dsh Unlimited OCR Skill

Aidenwu0209

Unlimited-OCR for DeepSeek Harness with a native tool and GUI configuration

Not checked1
D

Deepseek Harness File Upload Ocr Plugin

BYYY-eng

DeepSeek Harness 文件上传与本地 OCR 插件 | File upload and local OCR plugin for PDF, Word, Excel, PowerPoint, images, and text files.

Not checked1
V

Vision Toolkit For Dsh V0 1 Maybe

jmjmj009gt

Zero-dependency vision OCR/Q&A toolkit (CLI + local web GUI) for OpenAI-compatible VLMs: Zhipu GLM, Qwen, OpenAI, OpenRouter, SiliconFlow

Not checked1
E

Eagleeye Mcp

baimaomaomao556

EagleEye MCP — pixel-accurate visual toolbox for Agents (screenshot, measure, OCR, regression)

Not checked0
D

Dsh Tesseract Ocr

maxwell-feng

No description provided.

Not checked0
S

SnapShot

L-mimimi

一款 Windows 截图工具:截图 · 离线 OCR 文字识别 · 桌面置顶钉图,单文件绿色版,双击即用,无需安装、无需联网。

Not checked0
I

Image Analysis Skill

SKL-666666

图片结构化分析技能:双引擎OCR+形状/表格/图标/布局识别,让纯文本模型看懂图片

Not checked0
D

Dsh Eye

wenliang9527

No description provided.

Not checked0
D

Dsh Qwen Multimodal

wuwangmao

DSH bundle: Qwen multimodal bridge — vision (qwen3-vl), speech-to-text (qwen3-asr), text-to-image (qwen-image), for DeepSeek Harness

Not checked0
D

Dsh Plugin Vision

qizhen2021

No description provided.

Not checked0
D

Dsh Vision Suite

princefrogdida-ux

Windows-first vision suite with image understanding, OCR, screenshot diffing, and multi-provider routing for DeepSeek Harness.

Not checked0
D

Dsh Plugin Vision

tdf1995

Vision for text-only LLMs in DeepSeek Harness (DSH): describe images / OCR / VQA via free Gemini & GLM vision APIs

Not checked0
D

Dsh Visionary

zhuiyueya

Give text-only DeepSeek models eyes — a DeepSeek Harness plugin that transparently converts chat images into OCR text + vision-model descriptions before they reach the LLM. Configure vision backends (GLM-4V, Qwen-VL, Gemini, Ollama…) right in the Models settings page; multi-backend fallback chain, double-layer caching, no config files.

Not checked0
D

Dsh Vision Relay

junhongchashui

零修改、零切换的 DeepSeek Harness 视觉能力插件:纯文本模型粘贴即读图片,云端 + 本地 Ollama 双后端自动切换,ModLens v2 风格结构化证据输出。

Not checked0
D

Dsh Vision Skill

shajinhui

Give text-only AI agents eyes — clipboard, local images, and URLs via Gemini or OpenAI-compatible vision APIs

Not checked0
Q

Qwen Vision Mcp

Bigcheese18

Qwen3-VL vision MCP server: OCR / screenshot / chart understanding for text-only LLM clients (Claude Code, DeepSeek Harness, etc.)

Not checked0