47 #ocr DeepSeek Harness (DSH) Plugins
Every indexed repository carrying the GitHub topic ocr, sorted by stars. The tag is the author's own word for it — the install verdict on each card is ours.
Repositories tagged #ocr, most-starred first
Modlens
liustack
The first vision plugin for DeepSeek Harness, and the vision bridge for every text-only coding agent. Paste an image, get structured JSON evidence (OCR, layout, semantics). | 全网第一个 DeepSeek Harness 视觉插件,为 DeepSeek、GLM 等纯文本模型外挂视觉能力,粘贴图片即得结构化 JSON 证据(OCR、版面、语义)。
Dsh Vision Toolkit
Anionex
让纯文本模型更好地做视觉任务的DeepSeek Harness插件:带意图的图片问答、长截图 OCR、UI 还原等|DeepSeek Harness-native integration for agent-vision-toolkit: image Q&A, long-screenshot OCR, UI restoration, grounding, pixel diff, Artifacts, and Web UI.
Dsh Vision Complete
Yts1919
给 DeepSeek 补上「眼睛和耳朵」的多模态视觉插件:看图 / OCR / 物体检测 / 视频理解 / 语音转写 / 截图直读,一键安装(DSH 插件)。
Dsh Vision
linenxi-ctrl
为 DeepSeek Harness 增加外挂识图模型:圆形鲸鱼按钮、发送图片识图自动回传、模型自主截图+识图工具、多协议自动适配、小白一键安装(未装 Node.js 自动下载)
Dsh Vision Proxy
Flyvhidbwo
DeepSeek Harness 插件:DeepSeek 大脑 + 自动识图。GUI 附加图片自动经 OpenAI 兼容 VLM 转译成文字后交给 DeepSeek 作答;支持百炼/智谱/OpenRouter 等任意 OpenAI 兼容端点(默认 qwen3.7-flash),无 key 自动探测本地 Ollama(图片不出本机);安装时有一问式确认
Dsh Docs
Sqhao-O
Fully local document intelligence for DeepSeek Harness. Parse PDF, Office files, images, and scanned documents with offline OCR. | DeepSeek Harness 全本地文档智能插件,支持 PDF、Office、图片与离线 OCR
Dsh Windows Ocr
maxwell-feng
No description provided.
Dsh Image Bridge
kbpoyo
DSH 插件:让纯文本模型也能看图。Web 端直接粘贴图片即可发送,无需指定图片路径;模型自主调用视觉技能查看,多模态模型原生直通,零skill绑定。
Dsh Plugin Deepeye
Favio8
DeepEye vision plugin for DeepSeek Harness (DSH): image description, OCR, VQA, UI layout, and clipboard analysis.
Codex Eyes Hands
651002
专为 DeepSeek Harness 打造:把本机 Codex CLI 变成纯文本 AI agent 的眼睛和手——看图/读文件/画图/监督执行/双通道容灾
Dsh Tool Ocr
ferstar
本地 OCR 插件:让纯文本生成 LLM 也能读懂图片 | Local OCR plugin: give text-only generative LLMs the ability to read images
Dsh Deepseek Vision
Argonaut790
Image understanding, OCR, and persistent visual evidence for text-only DeepSeek Harness models
Dsh Vision Primitives
zouyuanqing
Native interactive visual-reasoning plugin for DeepSeek Harness: precise pixel grounding (SOM grid / zoom / annotate / measure / diff / color / OCR) + MiMo V2.5 multimodal backend, zero external MCP servers.
Dsh Attachment Formats
linkingoscar
Codex-style attachment formats for the DeepSeek Harness Web GUI: PDF text-layer extraction, Office text extraction, scanned-PDF OCR, long-document spill + index cards, image-to-PNG.
GrassVison
moduqishi
给纯文本大模型装上原生视觉:流式真实思考链 · 跨轮次无感重看 · 像素级证据与 SVG 图元 · OpenAI/Anthropic/Responses 三协议兼容 | Native vision for text-only LLMs: streaming real thinking chain, cross-turn re-view, pixel-level evidence & SVG primitives, OpenAI/Anthropic/Responses compatible.
DeepSeek Prism
YOGEMOW
为纯文本模型按需识图:DSH 零补丁 Cordis 插件(prism_see 工具 + 图片 VEP 降级 + 技能运行时注册)+ Codex Skill;多 Provider 视觉 API,VEP/1 低 Token 视觉证据包
Deepseek Omnimodal
good-boy4069
Open-source multimodal MCP plugin for text-only AI agents: recognize and generate images, video, and audio through Qwen/DashScope. Supports Codex, Claude Code, and DeepSeek Harness ecosystem.
DSH Plugins 4U
honghudavy-star
DSH 自建插件集合:微信桥接器 + GUI 微信入口补丁,一键安装
Shadow Vision
WardLu
Open-source MCP vision server that gives text-only LLMs and AI agents image understanding, OCR, visual analysis, UI inspection, and multimodal capabilities.
Free Vision Skill
niyongsheng
Local‑only vision skill for macOS 本地化识图技能
Locallens
uknowmyface
Local OCR for DeepSeek Harness — read text from screenshots on your Mac with Apple's Vision framework. No API key, no upload.
Ds Vision Plugin
Sorwcyra
Paste images into DeepSeek Harness with a four-model vision race, OCR, and an automatic text bridge.
Dsh PaddleOCR Skills
Aidenwu0209
PaddleOCR skills for DeepSeek Harness with native tools and GUI configuration
Glm4v Vision Mcp
ethanweave
GLM-4.6V 图像理解 MCP:识图/OCR/图表解析,原生接入 DeepSeek Harness(dsh-mcp-client),也兼容 Codex/Cline 等
Dsh Vision Bridge
Xieweikang123
Give a text-only dsh model eyes: pasted images recognized into text via an OpenAI-compatible vision endpoint.
Deepseek Harness
tylerbuilds
Local-first, safety-gated CLI and MCP harness for bounded DeepSeek batch and corpus workloads.
Dsh Mac Vision
Kevoyuan
On-device macOS OCR and Apple Vision for DeepSeek Harness — one native plugin with a bundled Skill.
Easy Vision
Koreyer
A DeepSeek Harness tool plugin that lets text-only agents "see" local images — auto-detects the real format and returns a detailed text description via any OpenAI-compatible vision model.
Doubao Vision Dsh
hawkongz
让纯文本模型通过桌面豆包看见聊天图片的 DeepSeek Harness 宿主插件(CDP 桥接,全预设生效,识别可取消)
Dsh Legal Dashboard
ByronLeeeee
Matter-aware legal workspace dashboard and document agent tools for DeepSeek Harness
Sidesight
ZhuXinAI
CLI-first vision sidecar for text-only coding agents. Analyze screenshots, diagrams, charts, UI diffs, and videos with OpenAI-compatible multimodal models.
Dsh Unlimited OCR Skill
Aidenwu0209
Unlimited-OCR for DeepSeek Harness with a native tool and GUI configuration
Deepseek Harness File Upload Ocr Plugin
BYYY-eng
DeepSeek Harness 文件上传与本地 OCR 插件 | File upload and local OCR plugin for PDF, Word, Excel, PowerPoint, images, and text files.
Vision Toolkit For Dsh V0 1 Maybe
jmjmj009gt
Zero-dependency vision OCR/Q&A toolkit (CLI + local web GUI) for OpenAI-compatible VLMs: Zhipu GLM, Qwen, OpenAI, OpenRouter, SiliconFlow
Eagleeye Mcp
baimaomaomao556
EagleEye MCP — pixel-accurate visual toolbox for Agents (screenshot, measure, OCR, regression)
Dsh Tesseract Ocr
maxwell-feng
No description provided.
SnapShot
L-mimimi
一款 Windows 截图工具:截图 · 离线 OCR 文字识别 · 桌面置顶钉图,单文件绿色版,双击即用,无需安装、无需联网。
Image Analysis Skill
SKL-666666
图片结构化分析技能:双引擎OCR+形状/表格/图标/布局识别,让纯文本模型看懂图片
Dsh Eye
wenliang9527
No description provided.
Dsh Qwen Multimodal
wuwangmao
DSH bundle: Qwen multimodal bridge — vision (qwen3-vl), speech-to-text (qwen3-asr), text-to-image (qwen-image), for DeepSeek Harness
Dsh Plugin Vision
qizhen2021
No description provided.
Dsh Vision Suite
princefrogdida-ux
Windows-first vision suite with image understanding, OCR, screenshot diffing, and multi-provider routing for DeepSeek Harness.
Dsh Plugin Vision
tdf1995
Vision for text-only LLMs in DeepSeek Harness (DSH): describe images / OCR / VQA via free Gemini & GLM vision APIs
Dsh Visionary
zhuiyueya
Give text-only DeepSeek models eyes — a DeepSeek Harness plugin that transparently converts chat images into OCR text + vision-model descriptions before they reach the LLM. Configure vision backends (GLM-4V, Qwen-VL, Gemini, Ollama…) right in the Models settings page; multi-backend fallback chain, double-layer caching, no config files.
Dsh Vision Relay
junhongchashui
零修改、零切换的 DeepSeek Harness 视觉能力插件:纯文本模型粘贴即读图片,云端 + 本地 Ollama 双后端自动切换,ModLens v2 风格结构化证据输出。
Dsh Vision Skill
shajinhui
Give text-only AI agents eyes — clipboard, local images, and URLs via Gemini or OpenAI-compatible vision APIs
Qwen Vision Mcp
Bigcheese18
Qwen3-VL vision MCP server: OCR / screenshot / chart understanding for text-only LLM clients (Claude Code, DeepSeek Harness, etc.)