Open-source SenseVoiceSmall model for Mandarin, Cantonese, English, Japanese, and Korean ASR, language ID, emotion recognition, and audio event detection.
-
Updated
Sep 10, 2026 - C
Open-source SenseVoiceSmall model for Mandarin, Cantonese, English, Japanese, and Korean ASR, language ID, emotion recognition, and audio event detection.
即時影音雙語字幕翻譯系統。本機離線 AI 語音辨識與翻譯,支援 SenseVoice、Ollama 與 DeepSeek。 | Real-time tab audio bilingual subtitle translation system using local ASR (SenseVoice) and LLMs (Ollama / DeepSeek). 100% offline & private.
Windows 离线语音输入:按 F2 说话,文字回到当前光标;不上传、可恢复、长语音不丢尾。
Voice-driven writing, input, and cross-app work for your desktop.
Free offline voice-to-text for macOS. Push-to-talk, works in any app.
Open-source offline voice dictation — a free alternative to Typeless. 100% local, zero internet, ideal for air-gapped & classified environments. SenseVoice + DirectML GPU, ~34MB, hotkey-driven. 纯离线语音输入,Typeless 免费平替,数据永不出设备,适合党政机关等涉密场景。
AI Speech Solutions for Tasks such as ASR, Vocal Extraction, Accompaniment Extraction, Audio Denoising, and Enhancement, Support models such as paraformer, sensevoice, fireredasr, zipformer, moonshine, wenet, whisper, fsmn-vad, silero-vad, CT Transformer punc, Spleeter, Uvr5, etc, apply ONNX models in various scenarios.
c# library for decoding paraformer, sensevoice Models,used in speech recognition (ASR)
A Rust-based, SenseVoiceSmall
Real-time bilingual subtitles for any browser video. Offline-first speech recognition (Whisper + SenseVoice, 100+ languages) with local Ollama AI translation — 100% private, ultra-low latency.
本地优先的 macOS 与 Windows 语音输入工具,支持端侧 SenseVoice 识别、全局语音输入、文件转写及可选 AI 校对。Local-first voice dictation for macOS and Windows with on-device SenseVoice, typing into any app, file transcription, and optional AI proofreading.
Tool for speech recognition using sensevoice-small
基于 SenseVoice 的 Windows 本地语音转文字工具,支持 OpenAI 格式 API 润色,低延迟,高精度。
中文语音助手 | 唤醒词 + ASR + OpenClaw Agent + TTS | 离线唤醒、流式语音交互、工具调用、Skills 扩展
Windows 微信 4.1.12+ 聊天记录本地导出工具:TXT / Markdown / JSON,支持图片、文件、语音、视频、本地语音转文字和聊天预览,面向 GPT、Claude、Gemini 等 LLM。 | Local-first WeChat chat exporter for LLMs.
本地优先 AI 语音输入工具 | Local-first AI voice typing tool. 支持 Whisper/SenseVoice/Parakeet/Qwen asr,离线可用,LLM 智能润色. Open-source speech-to-text with LLM post-processing, works offline.
Real-time speech-to-text clipboard tool with Silero VAD and local ASR support
Native local-first macOS speech-to-text, live captions, subtitle editing, offline translation, and on-device Gemma 4 transcript enhancement.
Windows 线上会议面试助手:多会议平台路由、本地流式语音识别、可选领域模块,并针对金融与会计场景做专项优化。
实时面试提词器:本地 ASR + BM25 检索 + LLM 快答,说完 1.3~1.5 秒出答案(Windows)
To associate your repository with the sensevoice topic, visit your repo's landing page and select "manage topics."