Solid free vision and image-generation skills for DeepSeek Harness with multi-provider failover and no committed keys.
Code quality7
README describes a structured Python bundle with a documented failover chain and bundled patch files for four Harness versions, but no tests or CI are mentioned in the provided excerpt.
Security7
The README explicitly states keys are never committed and images go only to user-chosen providers, though the install path pulls code from GitHub and the required Harness patches modify core behavior.
Utility9
Solves a concrete, common gap by letting text-only DeepSeek Harness sessions read pasted images and generate new ones via free models like GLM-4V-Flash and Kolors.
Maintenance8
Last updated 2026-08-24 with explicit support for rc.7, rc.8, v0.1.1-rc.1 and rc.2, and the repo is not archived, indicating active tracking of upstream releases.
Docs8
README is thorough with badges, comparison table, capability matrix, quick-start links and translations into nine languages, though the excerpt cuts off mid-table and full setup details live in linked docs.
Generated by AI after reading the project README, as a decision aid; neutral scores are given when information is thin. Not an official endorsement.
Free image reading & generation for DeepSeek Harness (rc.7 / rc.8 / v0.1.1-rc.1 / rc.2) — paste-image reading with auto vision transcription, DeepSeek-V4-Flash-Vision-Exp / GLM-4V-Flash / SenseNova / Gemini failover, Kol
How popular is dsh-media-skills?
dsh-media-skills has 17 stars and 2 forks on GitHub, last updated 2026-08-24.
What license does dsh-media-skills use?
dsh-media-skills is distributed under the MIT license.
Is dsh-media-skills compatible with DeepSeek Harness?
dsh-media-skills is listed in the DSH Universe directory as a tool for DeepSeek Harness. This plugin is listed in the DSH Universe directory and covered by its validation pipeline.
CLASSIFICATION EVIDENCE
/10类依据
项目类型技能
功能/10类文件与数据
规则置信度High
System reads GitHub Topics first, then compares against the in-site category dictionary and word-root rules. Current match:agent-skills、skill、image-generation、vision。
The first vision plugin for DeepSeek Harness, and the vision bridge for every text-only coding agent. Paste an image, get structured JSON evidence (OCR, layout, semantics). | The strongest vision add-on plugin for DeepSeek Harness, adding vision capability to text-only models like DeepSeek and GLM, paste an image and get structured JSON evidence (OCR, layout, semantics).
[dsh] A more powerful visual toolkit for text-only models: one-line install and use, paste images for direct recognition, multi-image Q&A, screenshot-to-frontend UI restoration, and more|DeepSeek Harness-native integration for agent-vision-toolkit: image Q&A, long-screenshot OCR, UI restoration, grounding, pixel diff, Artifacts, and Web UI.