Back to catalog

GITHUB TOPIC

vision

159projects include this topic

The "vision" topic on GitHub groups 159 open-source projects in the DeepSeek Harness (DSH) ecosystem, led by modlens with 4.1k GitHub stars. modlens — The first vision plugin for DeepSeek Harness, and the vision bridge for every text-only coding agent. Every project here is indexed by DSH Universe with live GitHub data — stars, activity and install status — so you can compare and install directly.

{count} projects

Exact GitHub Topic match

Files & dataSkill

modlens

liustack

The first vision plugin for DeepSeek Harness, and the vision bridge for every text-only coding agent. Paste an image, get structured JSON evidence (OCR, layout, semantics). | The strongest vision add-on plugin for DeepSeek Harness, adding vision capability to text-only models like DeepSeek and GLM, paste an image and get structured JSON evidence (OCR, layout, semantics).

Structure check pending
agent-skillsclaude-codeclaude-skillscodex
DSH PluginsSkill

A better vision toolbox and skill for pure-text models to "see": multi-image understanding, image Q&A, frontend UI restoration, GUI automation and more, with optional seamless integration into multiple mainstream agents, recognizing pasted images directly | A vision toolkit and skill designed for text-only llms — image Q&A, long-screenshot OCR, frontend UI restoration, and GUI automation, with optional seamless integration for Codex, Claude Code, Pi, Oh My Pi, and OpenCode

Structure check pending
agentagent-skillsclaude-codecodex
DSH PluginsPlugin

Eyes for text-only DeepSeek Harness agents: built-in free vision chain (no key) + pixel-level vision tools (Q&A, grounding, crop, pixel diff, colors, OCR, SVG trace, cutout, screenshots). One-command install, no Python, image turns work like ordinary tool-calling turns.

Structure check pending
deepseek-harnessdshdsh-pluginmultimodal
Files & dataPlugin

dsh-vision

oil-oil

Near-native image understanding for DeepSeek Harness

Structure check pending
deepseek-harnessdsh-pluginimage-understandingmultimodal
Files & dataPlugin

DSH plugin: pixel-to-text image reading for text-only models. image_scan/image_ocr/image_sample tools + image-reading skill (34-image trained methodology). Pure local, optional PaddleOCR.

Structure check pending
deepseek-harnessdsh-pluginimage-readingocr
Files & dataSkill

Free image reading & generation for DeepSeek Harness (rc.7 / rc.8 / v0.1.1-rc.1 / rc.2) — paste-image reading with auto vision transcription, DeepSeek-V4-Flash-Vision-Exp / GLM-4V-Flash / SenseNova / Gemini failover, Kolors + U1 Fast generation. No keys in repo.

Structure check pending
agent-skillsdeepseek-harnessdsh-pluginimage-generation
Files & dataPlugin

DeepSeek Harness plugin: DeepSeek Pro brain + automatic image recognition. Images attached in the GUI are processed by default with the official deepseek-v4-flash-vision-exp native vision model, converted to text, and passed to DeepSeek for answers (even text-only V4-Pro can see images); supports any OpenAI-compatible VLM such as Bailian/Zhipu/OpenRouter; auto-detects local Ollama without a key; one-question confirmation during install

Structure check pending
dashscopedeepseek-harnessdsh-pluginimage-understanding
Files & dataPlugin

Dsh-visual-plugin.Give your text-only model eyes: forward user images to any OpenAI-compatible vision model and see the results in a Web UI right panel

Structure check pending
deepseekdeepseek-harnessdsh-pluginimage-description
Files & dataPlugin

dsh-vision

linenxi-ctrl

Adds external image recognition models to DeepSeek Harness: round whale button, send image for recognition with auto-return, model autonomous screenshot + image recognition tools, automatic multi-protocol adaptation, one-click install for beginners (auto-download if Node.js is not installed)

Structure check pending
aideepseekdeepseek-harnessdsh
Files & dataPlugin

DSH plugin: images and files straight to text-only models — images keep the native attachment experience, PDF/Office/archives/video/audio show as square chips in the attachment rail, auto-converted to workspace paths on send; pairs with dsh-vision-toolkit for paste-to-view. A DSH plugin that delivers images AND files to text-only models as workspace paths: images keep the native attachment UI, other files show as square chips in the rail, paths append on send — pairs with dsh-vision-toolkit.

Structure check pending
agentattachmentsdeepseekdeepseek-harness
Files & dataPlugin

One tool = all MiniMax multimodal capabilities: DSH text-only models see images/draw images/generate video/speak/sing/cover/search/check quota | One mmx_bridge tool = all MiniMax multimodal (VLM/image/video/speech/music/cover/search/quota) for DeepSeek Harness (DSH)

Structure check pending
agent-toolai-agentcordisdeepseek-harness
Agents & sessionsPlugin

dsh-auxiliary

dsh-plugins

Auxiliary models for DeepSeek Harness: vision understanding and context compression through dedicated model routes. DeepSeek Harness auxiliary model plugin: provides independent model routes, tools and system prompts for vision understanding, context compression, approval review, sub-agents, session titles and image generation, never touching the main conversation model.

Structure check pending
approvalcontext-compressiondshdsh-plugin
Files & dataPlugin

A vision plugin built for DSH (DeepSeek Harness), now supports Agent-invoked image display / Vision plugin for DSH(DeepSeek Harness),support Proactive Image Display.

Structure check pending
deepseek-harnessdshdsh-pluginfree
Files & dataPlugin

On-demand vision for text-only DeepSeek Harness (DSH) sessions: images become markers, and a vision_describe tool sends only image + question to an OpenAI-compatible vision model

Structure check pending
agentdeepseekdeepseek-harnessdsh
Files & dataPlugin

DeepSeek Harness plugin: describe_image — give a text-only model vision through an OpenAI-compatible VLM endpoint

Structure check pending
deepseekdeepseek-harnessdescribe-imagedsh
Files & dataPlugin

DSH-AUX

DoloresCaritasAngelus

Auxiliary model system for DeepSeek Harness: unified aux-LLM routing (per-task model, timeout, concurrency, failure cooldown, main-model fallback) + vision_analyze / web_extract / compress_text tools, settings page, and session image lifecycle cleanup.

Structure check pending
auxiliary-modelcordisdeepseek-harnessdsh
Files & dataPlugin

This repository does not yet provide a project description.

Structure check pending
deepseekdeepseek-harnessdshdsh-plugin
Files & dataPlugin

DeepSeek Harness plugin that bridges session images to pluggable vision APIs while keeping DeepSeek as the primary model.

Structure check pending
deepseekdeepseek-aideepseek-harnessdeepseek-harness-desktop
Files & dataSkill

Local‑only vision skill for macOS, on-device image recognition skill dsh-plugin

Structure check pending
deepseek-harnessdsh-pluginocrskills
Files & dataPlugin

Paste images into DeepSeek Harness with a four-model vision race, OCR, and an automatic text bridge.

Structure check pending
deepseekdeepseek-harnessdsh-pluginimage-to-text
Files & dataPlugin

Vision for text-only LLMs in DeepSeek Harness (DSH): describe images / OCR / VQA via free Gemini & GLM vision APIs

Structure check pending
deepseek-harnessdshdsh-plugingemini
Files & dataPlugin

DeepSeek Harness usage dashboard with API balance, daily spend, external vision-call accounting, per-model stats, call logs, cache rate, TTFT, and CSV export.

Structure check pending
dashboarddeepseek-harnessdshdsh-plugin
Files & dataPlugin

HD image recognition enhancement plugin exclusively for deepseek-v4-flash-vision-exp: relaxes DSH image limits + highres_read chunked image recognition tool.

Structure check pending
deepseek-harnessdsh-pluginimage-recognitionvision
Files & dataPlugin

[Discontinued] DeepSeek Harness vision bridge plugin: the new Harness natively supports image recognition, please use the native capability, this repository is for historical reference only.

Structure check pending
agentaiattachmentchat
Files & dataPlugin

DeepSeek Harness plugin: lets models that can't see images, such as deepseek-v4-flash, handle chat images too, with a built-in image recognition tool. Install: dsh plugin --profile web add dsh-image-pathify

Structure check pending
deepseekdeepseek-harnessdshdsh-plugin
Files & dataPlugin

DeepEye vision plugin for DeepSeek Harness (DSH): image description, OCR, VQA, UI layout, and clipboard analysis.

Structure check pending
cordisdeepseek-harnessdshdsh-plugin
Files & dataPlugin

Local OCR plugin: lets text-only generative LLMs read images | Local OCR plugin: give text-only generative LLMs the ability to read images

Structure check pending
deepseek-harnessdsh-pluginocrplugin
Files & dataPlugin

see_image tool for DSH: lets text-only models 'see' images by routing them to a configurable vision model (default Zhipu GLM-4V-Flash, free)

Structure check pending
deepseekdeepseek-harnessdsh-pluginglm
Files & dataPlugin

Give DeepSeek a pair of eyes and a paintbrush: paste screenshots/images straight into the conversation, the GLM vision model first transcribes the image content precisely (error messages, code, UI preserved verbatim), then DeepSeek continues with your question —— all in the same turn, seamless throughout; when an illustration is needed, DeepSeek automatically calls the text-to-image backend and shows the image in the conversation.

Structure check pending
deepseek-harnessdshdsh-plugindsh-plugins
Development toolsPlugin

Control a HarmonyOS phone from the DeepSeek Harness web page and let the AI recognize content

Structure check pending
deepseek-harnessdeveloper-toolsdsh-pluginharmonyos
Files & dataPlugin

Vision sidecar for DeepSeek Harness: accept pasted images on text-only models.

Structure check pending
deepseek-harnessdsh-pluginvision
Files & dataPlugin

dsh-vision-hub

xing666173

DeepSeek Harness EAC vision suite: 15 pixel-level vision tools (enhanced dsh-tool-vision) + bridge inline preview + drag-and-drop file upload, single-endpoint driven, clean conversation, EAC native settings support

Structure check pending
ai-toolsdeepseek-harnessdsh-pluginllm
Files & dataPlugin

dsh-pseudo-vision

DDDFXYqiming

Local OCR, color-statistics, pixel-scan, and metadata bridge for text-only DeepSeek Harness models; no external vision API.

Structure check pending
deepseek-harnessdsh-pluginocrvision
Files & dataPlugin

dsh-xiapan-media

dongsheng123132

Native vision, gpt-image-2 and Seedance plugins for DeepSeek Harness via Xiapan Cloud

Structure check pending
deepseek-harnessdsh-pluginimage-generationmultimodal
Files & dataPlugin

dsh-sight

Fu3rte

Plug-in vision for text-only DeepSeek Harness (dsh) models: built-in free/cheap VLM presets + multi-image batch analysis

Structure check pending
deepseek-harnessdsh-pluginvisionvlm
Files & dataPlugin

Vision routing and image generation for DeepSeek Harness through a fixed Mix model.

Structure check pending
deepseek-harnessdsh-plugingpt-image-2image-generation
Models & MCPPlugin

dsh-open-eyes

hyper-dsh-plugins

A lightweight DeepSeek Harness vision delegation tool for text-only routes, with native OpenAI Responses, Chat Completions, and Anthropic Messages adapters.

Structure check pending
anthropicdeepseek-harnessdsh-pluginmultimodal
Files & dataPlugin

Config-only DeepSeek Harness bundle for OpenAI-compatible vision models.

Structure check pending
deepseek-harnessdshdsh-pluginmultimodal
Files & dataPlugin

Codex-style attachment formats for the DeepSeek Harness Web GUI: PDF text-layer extraction, Office text extraction, scanned-PDF OCR, long-document spill + index cards, image-to-PNG.

Structure check pending
attachmentcordisdeepseekdeepseek-harness
Files & dataPlugin

Standalone screen capture for DeepSeek Harness (dsh): browser hotkeys plus an agent-facing capture+read tool. Forked out of @liustack/modlens#48 (upstream declined the feature).

Structure check pending
deepseek-harnessdsh-pluginscreenshotvision
Files & dataPlugin

dsh-vision

reimu-create

DSH plugin: text-only models (e.g. DeepSeek-V4) automatically see images via a vision model. Official surface-replace, cache-friendly, human transcript untouched. Vision bridge for text-only models

Structure check pending
deepseek-harnessdsh-pluginmultimodalvision
Files & dataPlugin

SF Vision Bridge — eyes for text-only-model DeepSeek Harness.

Structure check pending
deepseek-harnessdsh-pluginstepfunvision
Files & dataPlugin

dsh-vision

Terry12138qy

DeepSeek Harness vision plugin: provides image recognition for models without native vision (Alibaba Cloud Bailian qwen3.5-omni-plus, auto-switches to Zhipu glm-4.6v-flash on failure). Ported and adapted from claude-vision-skill. | Vision tool for DeepSeek Harness

Structure check pending
deepseek-harnessdsh-pluginimage-descriptionvision
Files & dataSkill

visual-review

wang-bool

dsh plugin supporting image upload and image recognition. Turns the ds experience multimodal

Structure check pending
deepseek-harnessdshdsh-plugindsh-plugins
InterfacePlugin

dsh-ui-spec

yumimanji

DeepSeek Harness plugin: turn UI screenshots into structured, implementation-grade web frontend specs. Deterministic geometry (sharp) + optional vision-model semantics, merged into one JSON + Markdown spec.

Structure check pending
deepseek-harnessdesign-tokensdsh-pluginui
Files & dataPlugin

Silent vision bridge for DeepSeek Harness: route chat images to a fixed vision model, preserve UI originals, and reuse observations across compaction and restarts.

Structure check pending
deepseek-harnessdshdsh-pluginimage
Files & dataPlugin

dsh-vision

314857493

Free GLM vision for text-only DeepSeek Harness: paste images in the GUI (auto-transcribe route) + vision tool + skill

Structure check pending
deepseekdeepseek-harnessdsh-pluginvision
Files & dataPlugin

dsh-eye-vision

AlloyPlane

This repository does not provide a project description yet.

Structure check pending
aideepseek-harnessdshdsh-plugin
Files & dataPlugin

dsh-plugins

Bernardxu123

Official bundle-spec plugin collection for DeepSeek Harness (dsh): dsh-vision image viewing + dsh-sensenova-image generation + dsh-client-stats-decimal two-decimal stats

Structure check pending
ai-agentdeepseek-harnessdsh-pluginimage-generation
Files & dataPlugin

dsh-deepseek-vision

Cheng-cheng9669

DeepSeek web vision bridge plugin for DeepSeek Harness (DSH)

Structure check pending
deepseek-harnessdeepseek-visiondshdsh-plugin
Files & dataSkill

dsh-vision-skill

DDDFXYqiming

Vision skill plugin for DeepSeek Harness (image analysis and OCR)

Structure check pending
agent-skillsdeepseek-harnessdshdsh-plugin
Files & dataPlugin

dsh-image-mmx

fengs2021

Give DSH text models eyes: images are automatically recognized via mmx (MiniMax VLM), and the results are injected into the model context

Structure check pending
deepseek-harnessdsh-pluginvision
Files & dataPlugin

Local-first vision for DeepSeek Harness: structured JSON evidence (OCR/layout/semantics) from local VLMs (LM Studio/Ollama), zero API cost, images never leave your machine.

Structure check pending
deepseek-harnessdsh-pluginlocal-firstvision
Files & dataPlugin

dsh-tool-eyes

go-farther-and-farther

DeepSeek Harness (DSH) local vision eyes plugin: screen tool (screenshots/images handed to a local vision model for description) + ocr tool (Windows built-in OCR extracts text character by character). Zero cloud, zero GPU for OCR, images never leave the machine.

Structure check pending
deepseek-harnessdsh-pluginocrvision
Models & MCPPlugin

dsh-vision-bridge

GooDAnDReaDY

Universal vision bridge for DeepSeek Harness: attachments with native models, 40+ tools, PDF/OCR/diagrams.

Unrecognized
deepseek-harnessdshdsh-pluginmultimodal
CommunicationChannel integration

DSH_plugins_4U

honghudavy-star

DSH custom plugin collection: WeChat bridge + GUI WeChat entry patch, one-click install

Structure check pending
ai-agentchatbotdeepseek-harnessdsh
Files & dataPlugin

DeepSeek Harness (DSH) vision plugin: image recognition via Edge + Doubao Web, zero cost, no API Key. General recognition + math-modeling diagram specialty (geometry/flowcharts/charts/tables/formulas) + uncertainty clarification loop. Vision plugin for DeepSeek Harness: image understanding via Edge + Doubao Web, zero cost, no API key. General recognition + math-modeling diagrams (geometry/flowcharts/charts/tables/formulas) + clarify loop for uncertainties.

Structure check pending
deepseek-harnessdsh-pluginvision
Files & dataPlugin

DeepSeek Harness image recognition plugin: stays in the DeepSeek conversation, 15+ providers' vision models translate images into text, configurable in Settings→Plugins

Structure check pending
deepseek-harnessdsh-pluginvision
Files & dataPlugin

dsh-mindseye

kanchengw

Plug-in vision for text-only models on DSH, with native interaction for image understanding and generation, GUI automation, through layered evidence memory and cache.

Structure check pending
agentdeepseek-harnessdsh-plugingui-automation
Files & dataPlugin

A DeepSeek Harness tool plugin that lets text-only agents "see" local images — auto-detects the real format and returns a detailed text description via any OpenAI-compatible vision model.

Structure check pending
cordis-plugindeepseek-harnessdsh-pluginimage
Files & dataPlugin

dsh-eyes

Leeminjing

Give text-only DeepSeek models on-demand vision: upload images, DeepSeek answers by calling a view_image tool backed by any OpenAI-compatible vision endpoint (Qwen/DashScope by default).

Structure check pending
dashscopedeepseek-harnessdsh-pluginmultimodal
Files & dataPlugin

Transparent image preprocessing route for DeepSeek Harness

Structure check pending
ai-agentscordisdeepseekdeepseek-harness
Files & dataPlugin

Vision-augmented DeepSeek adapter plugin for DeepSeek Harness: a vision-capable model describes image input, then a text-only DeepSeek model reasons over the description

Structure check pending
deepseek-harnessdsh-pluginllmplugin
Files & dataPlugin

Give DeepSeek Harness text-only models vision: Codex-style drag-and-drop of images into the chat box automatically routes them to the user-configured vision model for text conversion, v4 reads images without switching models. Vision for text-only DSH models: routes chat-box images to a user-configured vision model and returns text descriptions, deepseek-v4-pro reads images without switching.

Structure check pending
codex-styledeepseek-harnessdsh-pluginimage
Files & dataPlugin

dsh-auto-vision

NormanFxxkingRockwell

DeepSeek Harness vision bridge: automatically discovers your configured multimodal models and equips text-only main models with a vision tool, returning results as plain text. Zero config, one-command install.

Structure check pending
cordicdeepseek-harnessdshdsh-plugin
Files & dataPlugin

DeepSeek VisionPlus — official-grade vision extension for DeepSeek Harness. Routes image understanding to a free vision-model pool (Zhipu GLM, SiliconFlow Qwen) with automatic fallback, rate limiting, one-click platform tests and live status lines; text stays on DeepSeek. One-command install. MIT.

Structure check pending
deepseekdeepseek-harnessdshdsh-plugin
Models & MCPPlugin

dsh-qwen-mm

RRRosmontis

Qwen-MM-Plugins integration bundle for DeepSeek Harness (dsh) — multimodal MCP tools (vision, OCR, ASR, search, video, Blender, FreeCAD) + image attachment bridge. Enables native multimodal support in DeepSeek Harness.

Structure check pending
agentaideepseekdeepseek-harness
Files & dataPlugin

Multimodal plugin for DeepSeek harness, close to a native experience.

Structure check pending
deepseek-harnessdoubaodshdsh-plugin
Models & MCPPlugin

text-llm-vision

shaoqiuyuavailable

Scene-aware vision routing layer for DeepSeek Harness (dsh): decides which engine/backend an image should go to (chat/UI/table/code) before other vision plugins route it. Switch-gated, front-loaded, never touches other plugins' tools. Scene-level vision routing layer.

Structure check pending
deepseek-harnessdshdsh-pluginlocal-ai
Files & dataPlugin

Bridges images into text for non-vision DeepSeek Harness models — your message stays untouched, zero manual setup.

Structure check pending
deepseek-harnessdsh-pluginvision
Files & dataPlugin

dsh-design-qa

sunxin-ai

Design-fidelity QA for DeepSeek Harness: lend any text-only model an eye, then judge whether the implementation matches the mock. Ships the benchmark behind that judgement — four fixtures, 23 injected defects, and every raw model transcript. Retires itself when DeepSeek ships vision.

Structure check pending
benchmarkdeepseek-harnessdesign-qadesign-review
Files & dataPlugin

dsh-browser-vision

tristan-mcinnis

Browser tool for DeepSeek Harness that can SEE the page: browser-use over CDP driven by deepseek-v4-flash-vision-exp. Reads canvas text, text inside images and rendered charts, returns schema-validated JSON, and reports per-run cost.

Structure check pending
browser-automationbrowser-usecdpdeepseek
Files & dataPlugin

mimo-vision

wulusai2333

DeepSeek Harness (DSH) native plugin — describe_image tool: a vision bridge (image → mimo-v2.5 → text description) over the ctx.fs / ctx.credentials seams

Structure check pending
agentcordisdeepseek-harnessdsh
Models & MCPPlugin

Eyes for text-only DeepSeek: view_image tool (local Ollama or any OpenAI-compatible VLM) + chat image-attachment bridge — paste/drop images in the chat and the model can see them.

Structure check pending
deepseek-harnessdshdsh-pluginmultimodal
Files & dataPlugin

DeepSeek Harness all-in-one: no model switching — regular DeepSeek auto-routes to vision & image gen. Multi-backend: Gemini + any OpenAI-compatible (GPT-4o, Qwen-VL, GLM-4V, gpt-image, DALL-E, Flux, OpenRouter). gemini_vision/gemini_generate_image/gemini_optimize_image with vision self-check. Better than modlens.

Structure check pending
aicordisdeepseekdeepseek-harness
Files & dataSkill

DeepSeek Harness (DSH) plugins. qwen-image gives a text-only coding model eyes: an image goes to a Qwen-VL route through ctx.llm and comes back as text, so DeepSeek keeps coding while Qwen looks. Pure ESM, no build permission at install. | DSH plugin set: qwen-image lets text-only models read images via Qwen VL and return text; pure ESM, no build authorization needed at install.

Structure check pending
agent-skillsclaude-codecodexcoding-agent
Files & dataPlugin

Native interactive visual-reasoning plugin for DeepSeek Harness: precise pixel grounding (SOM grid / zoom / annotate / measure / diff / color / OCR) + MiMo V2.5 multimodal backend, zero external MCP servers.

Structure check pending
ai-agentscomputer-visioncordisdeepseek-harness
Files & dataPlugin

dsh-llm-vision

1710782766

Reliable vision + OCR for text-only models on DeepSeek Harness: describe_image (normal/critical) + extract_text tools, auto-preprocessing, retries, and a persistent answer cache.

Structure check pending
deepseek-harnessdshdsh-pluginimage-understanding
Files & dataPlugin

Add image recognition to text-only DeepSeek Harness models: analyze_image forwards images to any OpenAI-compatible vision endpoint | Vision bridge for text-only DSH models

Structure check pending
deepseek-harnessdsh-pluginmultimodalopenai-compatible
Files & dataPlugin

c-vision

cczzyy-cn

DeepSeek Harness (DSH) vision plugin — gives the agent screen/window vision + computer-use ability (see/ocr/list_windows + mouse and keyboard control), cross-language calls to the bundled Python cvision, available on Windows / macOS.

Structure check pending
computer-usedeepseek-harnessdsh-pluginvision
Files & dataPlugin

DeepSeek Harness (DSH) Web plugin: drag-and-drop file preview / Markdown rendering / file box / smart desktop screenshot (fullscreen·region·window) / external image recognition and OCR text ingestion. Full development history with 21 snapshot tags.

Structure check pending
ai-agentdeepseek-harnessdrag-and-dropdsh
Files & dataPlugin

aura-vision

Ck-epsilon

Aura Vision - free vision OCR plugin for DeepSeek Harness web profile (permanent bundle, GLM-4V-Flash free tier, tile-based long-document recognition)

Structure check pending
deepseek-harnessdsh-pluginglm-4vocr
Files & dataPlugin

dsh-Sight

cransmathenia666-hash

Give text-only DeepSeek Harness (dsh) agents vision — pasted images auto-convert to text descriptions with persistent caching, each image converted only once.

Structure check pending
deepseek-harnessdshdsh-pluginimage
Files & dataSkill

dsh-picture-fit

cyanfish-x

DSH plugin + Agent Skill: auto-fit oversized images with sharp before attachment admission

Structure check pending
agent-skilldeepseekdeepseek-harnessdsh
Files & dataPlugin

pi-pseudo-vision

DDDFXYqiming

Local OCR + color-statistics + pixel-scan + metadata bridge for text-only Pi Coding Agent models. Pi port of dsh-pseudo-vision, no external vision API.

Structure check pending
deepseek-harnessimage-to-textlocal-firstmit-license
Files & dataPlugin

dsh-glm-vision

fightingFirefox

Connect Zhipu GLM vision models in dsh, letting text models like DeepSeek see images through the glm_vision tool.

Structure check pending
deepseek-harnessdsh-pluginglmmultimodal
Files & dataSkill

dsh-image-unlock

FrostLeafKEE

DeepSeek Harness plugin: lifts the Web GUI image input limit, image attachments are textualized and handed to a vision skill | dsh plugin that lifts the image-input gate and hands attachments to a vision skill

Structure check pending
deepseek-harnessdsh-pluginpluginskill
Files & dataPlugin

dsh-tool-read-tiff

fulander0301

Model-facing read_tiff tool for DeepSeek Harness: decodes TIFF/TIF images (multi-page, LZW/Deflate/PackBits/CCITT/JPEG compression, bilevel, 8/16-bit and float) into viewable PNGs with full header metadata, plus optional vision-model description.

Structure check pending
deepseek-harnessdsh-pluginimage-decodingimage-processing
Files & dataPlugin

dsh-img

gmleong

Give text-only models eyes: analyze_image tool for DeepSeek Harness, backed by free Chinese vision APIs (GLM-4V-Flash / Qwen-VL) or any OpenAI-compatible endpoint. dsh plugin that gives text-only models eyes.

Structure check pending
agentdeepseek-harnessdsh-pluginvision
Files & dataPlugin

dsh-vision-guard

good-boy4069

This repository does not yet provide a project description.

Structure check pending
deepseek-harnessdsh-pluginvision
Files & dataPlugin

Vision tools for DeepSeek Harness: OCR, chart extraction, UI review, comparison & image-to-code via OpenAI- or Anthropic-compatible endpoints, with a built-in FREE anonymous vision source and automatic rate-limit failover. Supports OpenAI/Anthropic compatible vision endpoints.

Structure check pending
deepseekdeepseek-harnessdsh-pluginfree
Files & dataPlugin

DeepSeek Harness host plugin that lets text-only models see chat images through desktop Doubao (CDP bridge, works with all presets, recognition cancellable)

Structure check pending
cordisdeepseekdeepseek-harnessdoubao
Files & dataChannel integration

Image message bridge for text-only models in DeepSeek Harness (dsh): image blocks → text placeholder + local path, vision via qwen script

Structure check pending
bridgecordisdeepseekdeepseek-harness
Files & dataPlugin

DeepSeek Harness vision enhancement plugin: hands images to an external vision model for analysis and outputs plain-text evidence with coordinate-based visual primitives, so non-multimodal text models can also understand images, screenshots and documents in conversation.

Structure check pending
deepseek-harnessdeepseek-harness-plugindshdsh-plugin
Files & dataPlugin

dsh-quicksight

Isanti2016

No project description provided for this repository yet.

Structure check pending
deepseek-harnessdsh-pluginocrplugin
Agents & sessionsPlugin

DSH plugin: dispatch image-recognition tasks to an opencode-go mimo-v2.5 subagent via system-prompt injection

Structure check pending
ai-agentcordisdeepseekdeepseek-harness
Files & dataPlugin

Zero-dependency vision OCR/Q&A toolkit (CLI + local web GUI) for OpenAI-compatible VLMs: Zhipu GLM, Qwen, OpenAI, OpenRouter, SiliconFlow

Structure check pending
deepseek-harnessdsh-pluginglmllm
Files & dataPlugin

dsh-autovision

Junkrat9527

dsh-autovision: paste an image into a text-only model composer and a configured multimodal model transcribes it to text automatically. Twin-provider auto-routing + agent-callable read-image tool. No built-in keys, no relay.

Structure check pending
deepseek-harnessdsh-pluginvision
Files & dataPlugin

dsh-vision

kaaaaahn

DSH local vision capability plugin: macOS Vision OCR + ollama qwen3-vl semantic description + uploaded image bridge

Structure check pending
deepseek-harnessdsh-pluginlocal-aiocr
Files & dataPlugin

dsh-mingmu

Lab-sku

Mingmu VisionBridge - in-house vision bridge: when a blind model receives an image, it automatically calls a vision model to recognize it

Structure check pending
ai-plugindeepseekdeepseek-harnessdsh-plugin
Files & dataPlugin

dsh-pro-vision

lasdrder0705

DSH plugin: let DeepSeek-V4-Pro use V4-Flash-Vision-Exp for attached images. Install: dsh plugin --profile web add github:lasdrder0705/dsh-pro-vision

Structure check pending
deepseek-harnessdsh-pluginvision
Files & dataPlugin

DeepSeek Harness vision plugin: analyze_image (structured OCR evidence) + capture_image (USB camera visual loop). Camera visual loop + structured evidence, supports Ollama / DeepSeek / Xiaomi three backends.

Structure check pending
cameradeepseek-harnessdsh-pluginimage-to-text
Files & dataPlugin

DSH vision bridge plugin: lets a non-vision main model see images (session image intake + auto transcription + view_image tool)

Structure check pending
deepseek-harnessdshdsh-pluginvision
Files & dataPlugin

A plugin that gives vision-less LLMs vision capabilities (via an external vision model, of course)

Structure check pending
deepseek-harnessdsh-pluginvision
DSH PluginsPlugin

DSH 投屏控制的共用核心:界面、设备列表、投屏面板与 AI 工具(配合其他 provider 使用)

Structure check pending
deepseek-harnessdshdsh-pluginscrcpy
Files & dataPlugin

A native DeepSeek Harness (DSH) Cordis plugin that analyzes images through the reverse-engineered chat.deepseek.com vision mode (model_type=vision) — free, no third-party vision API key required. Native DeepSeek Harness (DSH) Cordis plugin: analyzes images via the reverse-engineered chat.deepseek.com vision mode (model_type=vision) — free, no third-party vision API key required.

Structure check pending
deepseek-harnessdeepseek-harness-plugindeepseek-visiondsh-plugin
Files & dataPlugin

Automatic model routing for DeepSeek Harness: three difficulty tiers (hard/normal/easy) plus vision routing, picking models you already configured under Settings ? Models. ???? + ?????????

Structure check pending
classifierdeepseek-harnessdshdsh-plugin
Files & dataPlugin

Vision for DeepSeek Harness agents — paste images in the Web composer, delegate reads to Kimi/MiniMax vision routes on isolated contexts; zero image bytes in the main session

Structure check pending
ai-agentcordisdeepseekdeepseek-harness
DSH PluginsPlugin

Official enhancement suite for DeepSeek Harness — Vision, Soul/Persona, Long-term Memory & Plugin Marketplace.

Structure check pending
ai-agentdeepseek-harnessdshdsh-plugin
Files & dataPlugin

Native-vision Windows computer-use for DeepSeek Harness: screenshots, UIA, OCR, and approval-gated input

Structure check pending
automationcomputer-usedeepseek-harnessdeepseek-v4
Files & dataPlugin

dsh-vision

sjakdhasdh

Vision tool plugin for DeepSeek Harness (DSH): give text-only models like deepseek-v4-flash image recognition via Alibaba Bailian / any OpenAI-compatible vision API. Add an image recognition tool to DeepSeek Harness models without vision.

Structure check pending
ai-agentdeepseekdeepseek-harnessdsh-plugin
Files & dataPlugin

dsh-vision-link

sprainJinyu

Route-preserving image understanding for text-only models in DeepSeek Harness (DSH).

Structure check pending
deepseek-harnessdsh-pluginimage-understandingjavascript
Models & MCPPlugin

A see_image vision tool plugin for DeepSeek Harness — describe images through any OpenAI-compatible vision model (GitHub Copilot, OpenAI, Ollama, vLLM, LM Studio).

Structure check pending
deepseek-harnessdshdsh-plugingithub-copilot
Files & dataPlugin

This repository does not yet provide a project description.

Structure check pending
deepseek-harnessdeepseek-harness-plugindshdsh-plugin
Files & dataPlugin

dsh-vision-bridge

TwistedRiCen

DSH-native Vision Evidence bridge for text-only reasoning models with native image attachments and strict multi-image validation.

Structure check pending
deepseek-harnessdsh-pluginllmmultimodal
Files & dataPlugin

dsh-eye

wenliang9527

No project description provided for this repository yet.

Structure check pending
deepseek-harnessdsh-pluginocrvision
Files & dataPlugin

dsh-agnes-omni

wumu1111111

Agnes omni-modal plugin for DeepSeek Harness: agnes_vision (image understanding) + agnes_image (text-to-image / image-to-image) + a vision bridge that lets you send images in chat. API key via DSH credentials, never in code.

Structure check pending
agnesdeepseek-harnessdshdsh-plugin
Files & dataPlugin

dsh-photo-pick

xiaoyaoPanPan

DeepSeek Harness (dsh) plugin: rank similar photos with vision scoring. Install dsh-photo-pick-app.

Structure check pending
deepseekdeepseek-harnessdshdsh-plugin
Files & dataPlugin

Image recognition plugin for DeepSeek Harness: automatically detects the current model's vision capability, supports multi-provider vision model management and detection

Structure check pending
computer-visioncordis-plugindeepseek-harnessdsh-plugin
Files & dataPlugin

dsh-vision-bridge

Xieweikang123

Give a text-only dsh model eyes: pasted images recognized into text via an OpenAI-compatible vision endpoint.

Structure check pending
deepseek-harnessdsh-pluginimageocr
Files & dataPlugin

computer-use-vision

xuanyuanluoxue

Windows computer-use capability for DSH: screenshot, vision, simulated input, self-evolving knowledge base. Plugin + skill dual-mode.

Structure check pending
automationcomputer-usedeepseek-harnessdsh-plugin
Files & dataPlugin

dsh-vision

xzyonline

Vision for text-only DeepSeek: view_image tool via any OpenAI-compatible VLM endpoint. macOS/Windows/Linux, one-click install.

Structure check pending
ai-assisteddeepseekdeepseek-harnessdsh
Models & MCPPlugin

DeepSeek Harness vision completion: twin routing unlocks a native image experience, local Ollama request layer for image viewing, zero cloud dependency. Vision twin + local agentic vision tools for DeepSeek Harness.

Structure check pending
deepseek-harnessdsh-pluginollamavision
Files & dataPlugin

Vision toolkit for DeepSeek Harness -- give text-only agents eyes

Structure check pending
deepseek-harnessdeepseek-vldsh-pluginmultimodal
Files & dataPlugin

deepseek-visual-plugin

zhangzhimou78-code

dsh-plugin

Structure check pending
cordiscordis-plugindeepseekdeepseek-harness
Files & dataPlugin

dsh-vision

zoahdev

Give DeepSeek Harness eyes: analyze images with an OpenAI-compatible vision model via a vision_analyze tool.

Structure check pending
agentdeepseek-harnessdsh-pluginimage
Files & dataPlugin

Persistent vision sub-agent plugin: auto conversion of pasted images, describe_image tool, per-session isolated conversation memory with status capsule + detail page (DeepSeek Harness)

Structure check pending
deepseek-harnessdsh-pluginimagevision
Files & dataPlugin

Dynamic multimodal-to-text projection and cross-model vision bridge for DeepSeek Harness (dsh)

Structure check pending
claudecordisdeepseekdeepseek-harness
Files & dataPlugin

dsh-plugin-glm-vision

baldovinmarques391-design

GLM Vision plugin for DSH: image translation for non-multimodal models via GLM-4V-Flash

Structure check pending
deepseek-harnessdsh-pluginglmmultimodal
Files & dataPlugin

DSH plugin: auto-detect and configure model capabilities (reasoningEfforts + input modalities) for llm-pi-ai. Successor to dsh-reasoning-efforts.

Structure check pending
deepseek-harnessdshdsh-plugininput-modalities
Files & dataPlugin

This repository does not yet provide a project description.

Structure check pending
deepseek-harnessdshdsh-pluginmodel-router
Files & dataPlugin

Unified vision and image-generation plugin suite for DeepSeek Harness

Structure check pending
deepseek-harnessdsh-pluginimage-generationmultimodal
Files & dataPlugin

👁️ Give your DeepSeek Harness the gift of sight — enables pure-text LLMs to analyze images via Zhipu's free GLM-4V-Flash vision model

Structure check pending
deepseek-harnessdsh-pluginglm-4v-flashimage-analysis
Files & dataPlugin

dsh-vision-bridge

cyh12345678910

Multi-backend vision plugin for DeepSeek Harness — API (OpenAI Vision) + CDP (Doubao bridge), cross-platform, cached, configurable

Structure check pending
cdpcordisdeepseek-harnessdsh-plugin
Files & dataPlugin

Image input fallback for DeepSeek Harness with native multimodal model detection and Volcengine Ark vision.

Structure check pending
arkdeepseekdeepseek-harnessdsh-plugin
Files & dataPlugin

DeepSeek Harness multimodal vision bridge (dsh-vision-bridge): pasted images auto-converted to VL text descriptions via llm/stream (solves UNSUPPORTED_CONTENT) + view_image/ocr_image active vision tools + native multimodal routing auto-skip (rc.7); zero dependencies. Vision bridge for text-only DeepSeek models.

Structure check pending
dashscopedeepseek-harnessdshdsh-plugin
Files & dataPlugin

dsh-zhipu

fineven

Zhipu BigModel all-in-one plugin for the DeepSeek Harness: web search (web_search_prime) + web reader (webReader, server-side rendered) + GLM-4.6V vision (vision_analyze + pasted-image hook), one ZAI_API_KEY

Structure check pending
deepseek-harnessdshdsh-pluginglm
Files & dataPlugin

Give DeepSeek Harness eyes: one-click install the dsh-plugin-deepeye vision plugin with the free Zhipu GLM-4V-Flash model. Paste images into text-only LLMs

Structure check pending
deepeyesdeepseek-harnessdshdsh-plugin
Models & MCPPlugin

A zero-config, multi-provider vision tool for DeepSeek Harness with automatic local model discovery and privacy-aware remote fallback.

Structure check pending
deepseek-harnessdsh-pluginmultimodalollama
Files & dataPlugin

Bridge Apple on-device Vision framework (macOS) into DeepSeek Harness: OCR, image classification, face detection, document layout as local dsh tools. No network, no API key.

Structure check pending
deepseek-harnessdsh-pluginmacosocr
Files & dataPlugin

DSH plugin: image_analyze tool that sends screenshots/photos to an OpenAI-compatible vision model

Structure check pending
deepseek-harnessdsh-pluginvision
Files & dataPlugin

dsh-vision-relay

junhongchashui

Zero-modification, zero-switching vision plugin for DeepSeek Harness: text-only models read images on paste, cloud + local Ollama dual backends auto-switch, ModLens v2-style structured evidence output.

Structure check pending
cordisdeepseek-harnessdshdsh-plugin
Files & dataPlugin

DSH plugin: keep text-only DeepSeek models (V4-Flash / V4-Pro) and auto-route image-bearing requests to the official vision model (deepseek-v4-flash-vision-exp) - no manual model switching.

Structure check pending
deepseekdeepseek-harnessdshdsh-plugin
Files & dataPlugin

DeepSeek Harness (dsh) bundle plugin: vision_analyze tool lets text-only LLM agents read images via SenseNova VLM, with Schemastery config and single-source credentials

Structure check pending
deepseek-harnessdshdsh-pluginllm-plugin
Files & dataPlugin

modlens

Minglink

DeepSeek Harness external vision multimodal and OCR bridge plugin

Structure check pending
deepseekdeepseek-harnessdsh-pluginocr
Files & dataPlugin

This repo does not yet provide a project description.

Structure check pending
deepseek-harnessdshdsh-pluginocr
Files & dataPlugin

DSH Plugin: provides `doubao_ask` dynamic search/image generation/multimodal image recognition tools via local Quicker forwarding (doubao web2api), and supports **paste image → local path** (paste-to-path).

Structure check pending
cordisdeepseek-harnessdoubaodsh-plugin
DSH PluginsPlugin

Control macOS apps and browsers via text commands without touching your pointer.

Structure check pending
agent-toolsautomationbrowser-automationcross-platform
Files & dataPlugin

This repository does not yet provide a project description.

Structure check pending
bundledeepseek-harnessdshdsh-plugin
Files & dataPlugin

This repository does not yet provide a project description.

Structure check pending
ascii-artdeepseek-harnessdsh-pluginocr
Files & dataPlugin

Zero-core-change vision capability for DeepSeek Harness: the describe_image tool + profile bundle, installable via 'dsh plugin add'

Structure check pending
ai-agentsdeepseek-harnessdshdsh-plugin
Files & dataPlugin

This repository does not yet provide a project description.

Structure check pending
deepseekdeepseek-harnessdshdsh-plugin
Files & dataPlugin

DSH native plugin: structured image analysis via multimodal APIs (read_image_mimo tool) with provider failover, caching, SSRF protection and a Web UI config card; returns structured JSON evidence.

Structure check pending
deepseek-harnessdsh-pluginocrvision
Files & dataPlugin

dsh-ocr-bridge

vuvanmai936-dot

OCR-level vision bridge for DeepSeek Harness: paste images, read them locally (macOS Vision / Tesseract), answer with your text-only DeepSeek model

Structure check pending
deepseek-harnessdsh-pluginocrvision
CommunicationChannel integration

mydsh

wowayou

Personal Agent System on DeepSeek Harness — everything is a plugin: completion notifications, vision for text models, reply annotations, multi-session tabs, video support, sandbox patch

Structure check pending
agentdeepseek-harnessdshdsh-plugin
Files & dataPlugin

DSH bundle: Qwen multimodal bridge — vision (qwen3-vl), speech-to-text (qwen3-asr), text-to-image (qwen-image), for DeepSeek Harness

Structure check pending
asrdashscopedeepseek-harnessdsh-plugin
Files & dataPlugin

DeepSeek Harness plugin: upload/paste images; on send transcribe via a vision model (DashScope) or offline Windows OCR. Built with deepseek-harness, Made with the DeepSeek-harness.

Structure check pending
deepseek-harnessdshdsh-pluginvision
Files & dataPlugin

DeepSeek Harness plugin: route image-bearing messages to a user-configured OpenAI-compatible vision endpoint when the active text model cannot see images

Structure check pending
deepseek-harnessdsh-pluginvision
Files & dataPlugin

Let text-only models read images in DeepSeek Harness

Structure check pending
deepseek-harnessdshdsh-pluginimage-captioning