Skill_Seekers
yusufkaraaslan
Convert documentation websites, GitHub repositories, and PDFs into Claude AI skills with automatic conflict detection
GITHUB TOPIC
86projects include this topic
The "ocr" topic on GitHub groups 86 open-source projects in the DeepSeek Harness (DSH) ecosystem, led by Skill_Seekers with 15k GitHub stars. Skill_Seekers — Convert documentation websites, GitHub repositories, and PDFs into Claude AI skills with automatic conflict detection. Every project here is indexed by DSH Universe with live GitHub data — stars, activity and install status — so you can compare and install directly.
{count} projects
Exact GitHub Topic match
yusufkaraaslan
Convert documentation websites, GitHub repositories, and PDFs into Claude AI skills with automatic conflict detection
liustack
The first vision plugin for DeepSeek Harness, and the vision bridge for every text-only coding agent. Paste an image, get structured JSON evidence (OCR, layout, semantics). | The strongest vision add-on plugin for DeepSeek Harness, adding vision capability to text-only models like DeepSeek and GLM, paste an image and get structured JSON evidence (OCR, layout, semantics).
Anionex
[dsh] A more powerful visual toolkit for text-only models: one-line install and use, paste images for direct recognition, multi-image Q&A, screenshot-to-frontend UI restoration, and more|DeepSeek Harness-native integration for agent-vision-toolkit: image Q&A, long-screenshot OCR, UI restoration, grounding, pixel diff, Artifacts, and Web UI.
EthanYoQ
E-invoice organizing and reimbursement prep tool: batch-collect PDF/OFD/XML invoices from email, OCR, categorize and archive, and generate Excel summaries; offers Windows/macOS desktop versions and a DSH plugin.
mrpulor-gh
Desktop automation MCP server — computer use for any AI agent: control screen, windows, mouse/keyboard, and Chrome via Model Context Protocol (stdio). Not a DSH plugin — DSH users install dsh-nuphus-mcp.
tomwong001
情圣 · Claude Code 中文恋爱教练技能 · 微信/探探/Soul/Bumble/青藤之恋聊天截图分析 · 高情商回复生成 · 7阶段关系推进
jing-hy
DSH plugin: pixel-to-text image reading for text-only models. image_scan/image_ocr/image_sample tools + image-reading skill (34-image trained methodology). Pure local, optional PaddleOCR.
Flyvhidbwo
DeepSeek Harness plugin: DeepSeek Pro brain + automatic image recognition. Images attached in the GUI are processed by default with the official deepseek-v4-flash-vision-exp native vision model, converted to text, and passed to DeepSeek for answers (even text-only V4-Pro can see images); supports any OpenAI-compatible VLM such as Bailian/Zhipu/OpenRouter; auto-detects local Ollama without a key; one-question confirmation during install
Sqhao-O
Fully local document intelligence for DeepSeek Harness. Parse PDF, Office files, images, and scanned documents with offline OCR. | DeepSeek Harness fully local document intelligence plugin, supports PDF, Office, images, and offline OCR
linenxi-ctrl
Adds external image recognition models to DeepSeek Harness: round whale button, send image for recognition with auto-return, model autonomous screenshot + image recognition tools, automatic multi-protocol adaptation, one-click install for beginners (auto-download if Node.js is not installed)
maxwell-feng
This repository does not yet provide a project description.
balcoz
DeepSeek Harness local OCR plugin: paste an image, recognize text with PP-OCRv5 + ONNX Runtime, fully offline, supports TUI and Web | Local OCR plugin for DeepSeek Harness — paste an image, get its text via PP-OCRv5 + ONNX Runtime, fully offline. Supports TUI and Web.
sfyyy
On-demand vision for text-only DeepSeek Harness (DSH) sessions: images become markers, and a vision_describe tool sends only image + question to an OpenAI-compatible vision model
GOU-GEE
This repository does not yet provide a project description.
niyongsheng
Local‑only vision skill for macOS, on-device image recognition skill dsh-plugin
Sorwcyra
Paste images into DeepSeek Harness with a four-model vision race, OCR, and an automatic text bridge.
tdf1995
Vision for text-only LLMs in DeepSeek Harness (DSH): describe images / OCR / VQA via free Gemini & GLM vision APIs
Yurzi
Provider-independent DSH PDF parsing tools powered by MinerU.
Aidenwu0209
PaddleOCR skills for DeepSeek Harness with native tools and GUI configuration
Favio8
DeepEye vision plugin for DeepSeek Harness (DSH): image description, OCR, VQA, UI layout, and clipboard analysis.
ferstar
Local OCR plugin: lets text-only generative LLMs read images | Local OCR plugin: give text-only generative LLMs the ability to read images
xing666173
DeepSeek Harness EAC vision suite: 15 pixel-level vision tools (enhanced dsh-tool-vision) + bridge inline preview + drag-and-drop file upload, single-endpoint driven, clean conversation, EAC native settings support
beijingwahw
Vision-only desktop automation agent plugin for DeepSeek Harness (DSH) | Vision-only desktop automation Agent plugin: SoM grounding · Planner-Actor · effect verification · skill library
ByronLeeeee
Matter-aware legal workspace dashboard and document agent tools for DeepSeek Harness
DDDFXYqiming
Local OCR, color-statistics, pixel-scan, and metadata bridge for text-only DeepSeek Harness models; no external vision API.
elangan1997-cmyk
本地生图工作台(Lovart 平价平替):开自己的 API 生图,画布排版+修图/去背景/OCR/转矢量,可编辑 PSD/AI 交付,PS/AI 图层级双向桥接 · DSH 插件 npm: canvas-workbench
linkingoscar
Codex-style attachment formats for the DeepSeek Harness Web GUI: PDF text-layer extraction, Office text extraction, scanned-PDF OCR, long-document spill + index cards, image-to-PNG.
maxwell-feng
This repository does not yet provide a project description.
AlloyPlane
This repository does not provide a project description yet.
Argonaut790
Image understanding, OCR, and persistent visual evidence for text-only DeepSeek Harness models
DDDFXYqiming
Vision skill plugin for DeepSeek Harness (image analysis and OCR)
Fish121380
Windows desktop UI context picker for AI agents: select windows, UI elements, or screen regions with hover highlighting, UI Automation, screenshots, local OCR, and user-approved MCP output. Works with OpenAI Codex, DeepSeek Harness, and other MCP-compatible clients.
go-farther-and-farther
DeepSeek Harness (DSH) local vision eyes plugin: screen tool (screenshots/images handed to a local vision model for description) + ocr tool (Windows built-in OCR extracts text character by character). Zero cloud, zero GPU for OCR, images never leave the machine.
GooDAnDReaDY
Universal vision bridge for DeepSeek Harness: attachments with native models, 40+ tools, PDF/OCR/diagrams.
honghudavy-star
DSH custom plugin collection: WeChat bridge + GUI WeChat entry patch, one-click install
Koreyer
A DeepSeek Harness tool plugin that lets text-only agents "see" local images — auto-detects the real format and returns a detailed text description via any OpenAI-compatible vision model.
Leeminjing
Give text-only DeepSeek models on-demand vision: upload images, DeepSeek answers by calling a view_image tool backed by any OpenAI-compatible vision endpoint (Qwen/DashScope by default).
uknowmyface
Local OCR for DeepSeek Harness — read text from screenshots on your Mac with Apple's Vision framework. No API key, no upload.
xuxun-oss
DeepSeek Harness all-in-one: no model switching — regular DeepSeek auto-routes to vision & image gen. Multi-backend: Gemini + any OpenAI-compatible (GPT-4o, Qwen-VL, GLM-4V, gpt-image, DALL-E, Flux, OpenRouter). gemini_vision/gemini_generate_image/gemini_optimize_image with vision self-check. Better than modlens.
zjcdkj
DeepSeek Harness (DSH) plugins. qwen-image gives a text-only coding model eyes: an image goes to a Qwen-VL route through ctx.llm and comes back as text, so DeepSeek keeps coding while Qwen looks. Pure ESM, no build permission at install. | DSH plugin set: qwen-image lets text-only models read images via Qwen VL and return text; pure ESM, no build authorization needed at install.
zouyuanqing
Native interactive visual-reasoning plugin for DeepSeek Harness: precise pixel grounding (SOM grid / zoom / annotate / measure / diff / color / OCR) + MiMo V2.5 multimodal backend, zero external MCP servers.
zyh20041227
Full-coverage image tiling for DeepSeek Harness vision models, dense-text OCR, and document AI
1710782766
Reliable vision + OCR for text-only models on DeepSeek Harness: describe_image (normal/critical) + extract_text tools, auto-preprocessing, retries, and a persistent answer cache.
Aidenwu0209
Unlimited-OCR for DeepSeek Harness with a native tool and GUI configuration
Ck-epsilon
Aura Vision - free vision OCR plugin for DeepSeek Harness web profile (permanent bundle, GLM-4V-Flash free tier, tile-based long-document recognition)
DDDFXYqiming
Local OCR + color-statistics + pixel-scan + metadata bridge for text-only Pi Coding Agent models. Pi port of dsh-pseudo-vision, no external vision API.
Harvey-Will
Vision tools for DeepSeek Harness: OCR, chart extraction, UI review, comparison & image-to-code via OpenAI- or Anthropic-compatible endpoints, with a built-in FREE anonymous vision source and automatic rate-limit failover. Supports OpenAI/Anthropic compatible vision endpoints.
hawkongz
DeepSeek Harness host plugin that lets text-only models see chat images through desktop Doubao (CDP bridge, works with all presets, recognition cancellable)
henryxiao709
DSH-PDF plugin lets AI assistants read PDFs of any size: extract the full Unicode text layer via pdfjs-dist (Chinese, English and other writing systems), with automatic OCR for scanned/image pages — handwritten notes also become readable text. DSH-PDF plugin — read any-size PDFs in DeepSeek Harness: full Unicode text (Chinese/English) via pdfjs-dist + automatic OCR (Windows WinRT / tesseract.js) for scanned pages. MIT.
Isanti2016
No project description provided for this repository yet.
jmjmj009gt
Zero-dependency vision OCR/Q&A toolkit (CLI + local web GUI) for OpenAI-compatible VLMs: Zhipu GLM, Qwen, OpenAI, OpenRouter, SiliconFlow
kaaaaahn
DSH local vision capability plugin: macOS Vision OCR + ollama qwen3-vl semantic description + uploaded image bridge
Kevoyuan
On-device macOS OCR and Apple Vision for DeepSeek Harness — one native plugin with a bundled Skill.
leozou320-ai
Offline macOS Vision OCR for DeepSeek Harness — accurate, local, API-key free. | DeepSeek Harness local offline OCR plugin
lhbsaa
DeepSeek Harness vision plugin: analyze_image (structured OCR evidence) + capture_image (USB camera visual loop). Camera visual loop + structured evidence, supports Ollama / DeepSeek / Xiaomi three backends.
princefrogdida-ux
Windows-first vision suite with image understanding, OCR, screenshot diffing, and multi-provider routing for DeepSeek Harness.
PRTS168
Chat with, monitor, and approve your DSH (DeepSeek Harness) agents from WeChat over the clawbot iLink gateway: two-way text/images/voice/files/video, native vision or OCR, context-rotation policies, reminders, and a standalone admin console.
secretxuan
Native-vision Windows computer-use for DeepSeek Harness: screenshots, UIA, OCR, and approval-gated input
sprainJinyu
Route-preserving image understanding for text-only models in DeepSeek Harness (DSH).
wangzhanchao883
DeepSeek Harness plugin collection: self-developed DSH plugins (screenshot capture, OCR, Obsidian). ?? DSH ??????
wenliang9527
No project description provided for this repository yet.
xiaoyuink
Image recognition plugin for DeepSeek Harness: automatically detects the current model's vision capability, supports multi-provider vision model management and detection
Xieweikang123
Give a text-only dsh model eyes: pasted images recognized into text via an OpenAI-compatible vision endpoint.
cyh12345678910
Multi-backend vision plugin for DeepSeek Harness — API (OpenAI Vision) + CDP (Doubao bridge), cross-platform, cached, configurable
DreamRift
DeepSeek Harness multimodal vision bridge (dsh-vision-bridge): pasted images auto-converted to VL text descriptions via llm/stream (solves UNSUPPORTED_CONTENT) + view_image/ocr_image active vision tools + native multimodal routing auto-skip (rc.7); zero dependencies. Vision bridge for text-only DeepSeek models.
genusamblyrhynchusbrunooftoul602
Extend DeepSeek Harness composer to accept PDFs and more attachment formats Codex-style, with zero core changes and native pipeline reuse.
hamliy-feng
Visual model adapter for Codex and DeepSeek Harness, powered by PaddleOCR-VL and Qwen.
harmless0819-dev
DSH skill: read Chinese/ISO mechanical drawings into an evidence-anchored annotation document plus structured JSON (dimensions, tolerances, BOM).
Harzva
Bridge Apple on-device Vision framework (macOS) into DeepSeek Harness: OCR, image classification, face detection, document layout as local dsh tools. No network, no API key.
junhongchashui
Zero-modification, zero-switching vision plugin for DeepSeek Harness: text-only models read images on paste, cloud + local Ollama dual backends auto-switch, ModLens v2-style structured evidence output.
kid-tea
Local text-only OCR plugin for DeepSeek Harness: ocr_image tool extracts text from images locally — no vision model, no API key, no external upload.
L-mimimi
A Windows screenshot tool: capture · offline OCR text recognition · pin images on top of the desktop, single-file portable build, runs on double-click, no install, no internet.
Minglink
DeepSeek Harness external vision multimodal and OCR bridge plugin
paul-yangmy
This repo does not yet provide a project description.
pipiwolve
Baidu cloud OCR bundle for DeepSeek Harness: drag images/PDFs in, OCR to Markdown with PaddleOCR-VL or Unlimited-OCR. Baidu Cloud OCR plugin: drag in images/PDFs to recognize as Markdown and write to local files.
protoctistmoses143
Convert PDFs, Office docs, scanned images, and more to clean Markdown, JSON, or text locally with offline OCR—no servers, no API keys, fully private.
QEDQCD
This repository does not yet provide a project description.
qizhen2021
This repository does not yet provide a project description.
SKL-666666
Structured image analysis Skill: dual-engine OCR + shape/table/icon/layout recognition, letting text-only models understand images
Triple3h
DSH native plugin: structured image analysis via multimodal APIs (read_image_mimo tool) with provider failover, caching, SSRF protection and a Web UI config card; returns structured JSON evidence.
vuvanmai936-dot
OCR-level vision bridge for DeepSeek Harness: paste images, read them locally (macOS Vision / Tesseract), answer with your text-only DeepSeek model
wangzhanchao883
Point-and-shoot screenshot capture plugin for DeepSeek Harness: clipboard watcher + system floating window (comment & key-point, copy/save-doc/save-image) + instant OCR (Tongyi Qianwen) + Obsidian per-day merging + evening AI organization. Point-and-shoot · DSH screenshot-and-save plugin: clipboard watcher + system-level floating window + instant OCR + Obsidian daily merge + evening AI organization
wuwangmao
DSH bundle: Qwen multimodal bridge — vision (qwen3-vl), speech-to-text (qwen3-asr), text-to-image (qwen-image), for DeepSeek Harness
Ya-MiC
Zhanzhen — SME audit risk platform v1 framework (FastAPI + Vue3, local rule engine, evidence hash chain)
ydlstartx
AI-powered PDF reader for DeepSeek Harness with annotations, multi-PDF workflows, mixed image-text evidence, and on-demand OCR.
zhuiyueya
Give text-only DeepSeek models eyes — a DeepSeek Harness plugin that transparently converts chat images into OCR text + vision-model descriptions before they reach the LLM. Configure vision backends (GLM-4V, Qwen-VL, Gemini, Ollama…) right in the Models settings page; multi-backend fallback chain, double-layer caching, no config files.