dsh-omni-workstation
huashenglian
dsh全模态工作站插件,让模型支持视频、图片、语音的输入与输出,支持comfyui图像生成工具调用。Any-to-Any.
GITHUB TOPIC
6projects include this topic
The "speech-recognition" topic on GitHub groups 6 open-source projects in the DeepSeek Harness (DSH) ecosystem, led by dsh-omni-workstation with 3 GitHub stars. dsh-omni-workstation — dsh全模态工作站插件,让模型支持视频、图片、语音的输入与输出,支持comfyui图像生成工具调用. Every project here is indexed by DSH Universe with live GitHub data — stars, activity and install status — so you can compare and install directly.
{count} projects
Exact GitHub Topic match
huashenglian
dsh全模态工作站插件,让模型支持视频、图片、语音的输入与输出,支持comfyui图像生成工具调用。Any-to-Any.
huangdejie
Speech plugin for DeepSeek Harness: per-message voice playback, auto-announce, and dictation via mic — cloud TTS/ASR on Alibaba DashScope or Volcengine (freely combinable), with Web Speech API fallback.
leozou320-ai
Voice-to-text for the DeepSeek Harness Web UI — live, editable, never auto-sends. | DeepSeek Harness web voice input
wepar1212
DSH Web Chinese and English voice input plugin, supporting mixed Chinese-English dictation, built on the browser Web Speech API. Chinese and English voice input for DSH Web, with mixed-language dictation via the browser Web Speech API.
liyixuan201211
让 AI Agent 听懂音频:转写、鸟种识别、语音情感、环境声与音乐分析。全部本地运行,不用大模型。DeepSeek Harness 技能。
wkfedor
Local voice typing and speech-to-text plugin for DeepSeek Harness (dsh), powered by multilingual Whisper.