VultrVultr
返回目录

GITHUB TOPIC

ocr

43个项目包含此标签

43 个项目

GitHub Topic 精确匹配

文件与数据技能

让纯文本模型更好地做视觉任务的DeepSeek Harness插件:带意图的图片问答、长截图 OCR、UI 还原等|DeepSeek Harness-native integration for agent-vision-toolkit: image Q&A, long-screenshot OCR, UI restoration, grounding, pixel diff, Artifacts, and Web UI.

待结构检查
agent-skillsagent-vision-toolkitcomputer-visiondeepseek
文件与数据插件

dsh-vision

linenxi-ctrl

为 DeepSeek Harness 增加外挂识图模型:圆形鲸鱼按钮、发送图片识图自动回传、模型自主截图+识图工具、多协议自动适配、小白一键安装(未装 Node.js 自动下载)

待结构检查
aideepseekdeepseek-harnessdsh
文件与数据插件

DeepSeek Harness 插件:DeepSeek 大脑 + 自动识图。GUI 附加图片自动经 OpenAI 兼容 VLM 转译成文字后交给 DeepSeek 作答;支持百炼/智谱/OpenRouter 等任意 OpenAI 兼容端点(默认 qwen3.7-flash),无 key 自动探测本地 Ollama(图片不出本机);安装时有一问式确认

待结构检查
dashscopedeepseek-harnessdsh-pluginimage-understanding
文件与数据插件

dsh-docs

Sqhao-O

Fully local document intelligence for DeepSeek Harness. Parse PDF, Office files, images, and scanned documents with offline OCR. | DeepSeek Harness 全本地文档智能插件,支持 PDF、Office、图片与离线 OCR

待结构检查
ai-agentdeepseekdeepseek-harnessdocling
文件与数据插件

dsh-windows-ocr

maxwell-feng

该仓库暂未提供项目说明。

待结构检查
cordisdeepseek-harnessdshdsh-plugin
文件与数据技能

PaddleOCR skills for DeepSeek Harness with native tools and GUI configuration

待结构检查
agent-skillsdeepseek-harnessdshdsh-plugin
文件与数据插件

DeepEye vision plugin for DeepSeek Harness (DSH): image description, OCR, VQA, UI layout, and clipboard analysis.

待结构检查
cordisdeepseek-harnessdshdsh-plugin
文件与数据插件

该仓库暂未提供项目说明。

待结构检查
deepseekdeepseek-harnessdshdsh-plugin
文件与数据插件

DSH plugin: pixel-to-text image reading for text-only models. image_scan/image_ocr/image_sample tools + image-reading skill (34-image trained methodology). Pure local, optional PaddleOCR.

待结构检查
deepseek-harnessdsh-pluginimage-readingocr
文件与数据插件

Paste images into DeepSeek Harness with a four-model vision race, OCR, and an automatic text bridge.

待结构检查
deepseekdeepseek-harnessdsh-pluginimage-to-text
文件与数据插件

Image understanding, OCR, and persistent visual evidence for text-only DeepSeek Harness models

待结构检查
ai-agentscomputer-visiondeepseekdeepseek-harness
文件与数据插件

本地 OCR 插件:让纯文本生成 LLM 也能读懂图片 | Local OCR plugin: give text-only generative LLMs the ability to read images

待结构检查
deepseek-harnessdsh-pluginocrplugin
文件与数据插件

Codex-style attachment formats for the DeepSeek Harness Web GUI: PDF text-layer extraction, Office text extraction, scanned-PDF OCR, long-document spill + index cards, image-to-PNG.

待结构检查
attachmentdeepseek-harnessdshdsh-plugin
文件与数据技能

Local‑only vision skill for macOS 本地化识图技能

待结构检查
cordisdeepseek-harnessdsh-pluginocr
文件与数据插件

Native interactive visual-reasoning plugin for DeepSeek Harness: precise pixel grounding (SOM grid / zoom / annotate / measure / diff / color / OCR) + MiMo V2.5 multimodal backend, zero external MCP servers.

待结构检查
ai-agentscomputer-visioncordisdeepseek-harness
消息通讯渠道适配

DSH_plugins_4U

honghudavy-star

DSH 自建插件集合:微信桥接器 + GUI 微信入口补丁,一键安装

待结构检查
ai-agentchatbotdeepseek-harnessdsh
文件与数据插件

dsh-tesseract-ocr

maxwell-feng

该仓库暂未提供项目说明。

待结构检查
deepseek-harnessdshdsh-pluginlinux
文件与数据插件

Vision for text-only LLMs in DeepSeek Harness (DSH): describe images / OCR / VQA via free Gemini & GLM vision APIs

待结构检查
deepseek-harnessdshdsh-plugingemini
文件与数据插件

locallens

uknowmyface

Local OCR for DeepSeek Harness — read text from screenshots on your Mac with Apple's Vision framework. No API key, no upload.

待结构检查
ai-agentscordisdeepseekdeepseek-harness
文件与数据插件

dsh-vision-bridge

Xieweikang123

Give a text-only dsh model eyes: pasted images recognized into text via an OpenAI-compatible vision endpoint.

待结构检查
deepseek-harnessdsh-pluginimageocr
文件与数据插件

Unlimited-OCR for DeepSeek Harness with a native tool and GUI configuration

待结构检查
baidu-clouddeepseek-harnessdocument-parsingdsh-plugin
文件与数据插件

Matter-aware legal workspace dashboard and document agent tools for DeepSeek Harness

待结构检查
deepseek-harnessdocument-automationdsh-pluginlegal-ai
文件与数据技能

dsh-vision-skill

DDDFXYqiming

Vision skill plugin for DeepSeek Harness (image analysis and OCR)

待结构检查
agent-skillsdeepseek-harnessdshdsh-plugin
文件与数据插件

dsh-tool-eyes

go-farther-and-farther

DeepSeek Harness (DSH) 本地视觉眼睛插件:screen 工具(截图/图片交给本地视觉模型描述)+ ocr 工具(Windows 内置 OCR 逐字提取文字)。零云端、OCR 零 GPU、图片不出本机。

待结构检查
deepseek-harnessdsh-pluginocrvision
文件与数据插件

让纯文本模型通过桌面豆包看见聊天图片的 DeepSeek Harness 宿主插件(CDP 桥接,全预设生效,识别可取消)

待结构检查
cordisdeepseekdeepseek-harnessdoubao
文件与数据插件

Zero-dependency vision OCR/Q&A toolkit (CLI + local web GUI) for OpenAI-compatible VLMs: Zhipu GLM, Qwen, OpenAI, OpenRouter, SiliconFlow

待结构检查
deepseek-harnessdsh-pluginglmllm
文件与数据插件

On-device macOS OCR and Apple Vision for DeepSeek Harness — one native plugin with a bundled Skill.

待结构检查
apple-visioncomputer-visioncordisdeepseek
文件与数据插件

A DeepSeek Harness tool plugin that lets text-only agents "see" local images — auto-detects the real format and returns a detailed text description via any OpenAI-compatible vision model.

待结构检查
cordis-plugindeepseek-harnessdsh-pluginimage
文件与数据插件

dsh-eyes

Leeminjing

Give text-only DeepSeek models on-demand vision: upload images, DeepSeek answers by calling a view_image tool backed by any OpenAI-compatible vision endpoint (Qwen/DashScope by default).

待结构检查
dashscopedeepseek-harnessdsh-pluginmultimodal
文件与数据插件

dsh-vision-suite

princefrogdida-ux

Windows-first vision suite with image understanding, OCR, screenshot diffing, and multi-provider routing for DeepSeek Harness.

待结构检查
computer-visiondeepseek-harnessdsh-pluginocr
文件与数据技能

DeepSeek Harness (DSH) plugins. qwen-image gives a text-only coding model eyes: an image goes to a Qwen-VL route through ctx.llm and comes back as text, so DeepSeek keeps coding while Qwen looks. Pure ESM, no build permission at install. | DSH 插件集:qwen-image 让纯文本模型借千问 VL 读图,返回文本;纯 ESM,安装无需构建授权。

待结构检查
agent-skillsclaude-codecodexcoding-agent
文件与数据插件

dsh-attachment-formats

genusamblyrhynchusbrunooftoul602

Extend DeepSeek Harness composer to accept PDFs and more attachment formats Codex-style, with zero core changes and native pipeline reuse.

待结构检查
attachmentdeepseek-harnessdshdsh-plugin
文件与数据索引目录

dsh-pdf

henryxiao709

DSH-PDF插件,让 AI 助手读取任意大小的 PDF 文件: 通过 pdfjs-dist 提取完整 Unicode 文本层(中文、英文及其它文字系统),并对扫描件/图片页自动 OCR, 手写笔记也能变成可读文本。DSH-PDF plugin — read any-size PDFs in DeepSeek Harness: full Unicode text (Chinese/English) via pdfjs-dist + automatic OCR (Windows WinRT / tesseract.js) for scanned pages. MIT.

非插件验证范围
cordisdeepseekdeepseek-harnessdsh
文件与数据插件

dsh-quicksight

Isanti2016

该仓库暂未提供项目说明。

待结构检查
deepseek-harnessdsh-pluginocrplugin
文件与数据插件

dsh-vision-relay

junhongchashui

零修改、零切换的 DeepSeek Harness 视觉能力插件:纯文本模型粘贴即读图片,云端 + 本地 Ollama 双后端自动切换,ModLens v2 风格结构化证据输出。

待结构检查
cordisdeepseek-harnessdshdsh-plugin
文件与数据插件

Local text-only OCR plugin for DeepSeek Harness: ocr_image tool extracts text from images locally — no vision model, no API key, no external upload.

待结构检查
deepseekdeepseek-harnessdsh-pluginlocal
文件与数据插件

SnapShot

L-mimimi

一款 Windows 截图工具:截图 · 离线 OCR 文字识别 · 桌面置顶钉图,单文件绿色版,双击即用,无需安装、无需联网。

待结构检查
csharpdeepseek-harnessdsh-pluginocr
文件与数据插件

Offline macOS Vision OCR for DeepSeek Harness — accurate, local, API-key free. | DeepSeek Harness 本地离线 OCR 插件

待结构检查
apple-visiondeepseekdeepseek-harnessdsh-plugin
文件与数据插件

该仓库暂未提供项目说明。

待结构检查
ascii-artdeepseek-harnessdsh-pluginocr
文件与数据技能

图片结构化分析技能:双引擎OCR+形状/表格/图标/布局识别,让纯文本模型看懂图片

待结构检查
agent-skillsdeepseek-harnessdsh-pluginimage-analysis
文件与数据插件

dsh-eye

wenliang9527

该仓库暂未提供项目说明。

待结构检查
deepseek-harnessdsh-pluginocrvision
文件与数据插件

DSH bundle: Qwen multimodal bridge — vision (qwen3-vl), speech-to-text (qwen3-asr), text-to-image (qwen-image), for DeepSeek Harness

待结构检查
asrdashscopedeepseek-harnessdsh-plugin
文件与数据插件

dsh-visionary

zhuiyueya

Give text-only DeepSeek models eyes — a DeepSeek Harness plugin that transparently converts chat images into OCR text + vision-model descriptions before they reach the LLM. Configure vision backends (GLM-4V, Qwen-VL, Gemini, Ollama…) right in the Models settings page; multi-backend fallback chain, double-layer caching, no config files.

待结构检查
deepseekdeepseek-harnessdsh-pluginimage-understanding