Command line interface for the built-in speech recognition and transcription capabilities in macOS.
-
Updated
Aug 30, 2026 - Objective-C
Command line interface for the built-in speech recognition and transcription capabilities in macOS.
💬 Fast, cross-platform CLI and GUI for batch transcription, translation, speaker annotation and subtitle generation using OpenAI’s Whisper on CPU, Nvidia GPU and Apple MLX.
一个可以让 AI 帮你把台本转换为带时间轴的字幕的工具。
OCTRA is a web-application for the orthographic transcription of audio files.
Offline macOS menu bar app for speech-to-text transcription using AI models (NVIDIA Parakeet, Whisper). 100% local & private.
🎵 Complete offline audio transcription system with speaker diarization using OpenAI Whisper and PyAnnote. Features automatic audio cleaning, precise timestamps, multiple output formats (JSON/TXT/Markdown), and support for 20+ audio formats. No external APIs required - works entirely offline.
Modular tool for Digital Humanities: IIIF downloader + Studio environment. Supports PDF import, hybrid OCR/HTR, side-by-side manual correction, and global library search. 🛠️📜
French audio transcription using gradio
WhisperVoice is a browser extension that converts speech to text in real-time using speech recognition APIs. It’s perfect for quick transcriptions, note-taking, and accessibility, supporting multiple languages and customizable settings for a tailored experience.
Speakr — Free open source alternative to Wispr Flow. Desktop voice dictation with global hotkey, Groq Whisper transcription, and auto-paste into any app.
The Whisper Subtitle Generator leverages OpenAI's Whisper model to generate subtitles from audio and video files. This Python-based tool supports multiple languages and employs advanced audio processing techniques to ensure high accuracy in transcription.
AI-Video-Transcriber is an intelligent, open-source tool that automatically transcribes video and audio files using advanced artificial intelligence. It supports multiple languages, accurate speech recognition, and provides easy-to-read text transcripts for content creators, educators, and businesses.
Self-hosted AI tool to transcribe and summarize meetings. Upload audio files, transcribe with Whisper, and generate structured summaries using OpenAI GPT or Google Gemini.
Dictator – Supercharge Cursor Chat with voice-to-text, custom AI prompts, and workflow automation. Speak your ideas, inject templates instantly, and code faster with AI-powered assistance.
Audio transcription and practice tool with waveform visualization, loop regions, timeline tags, and pitch-preserving playback for musicians.
Open Video Transcribe - Open-source video transcription tool that emphasizes the primary use case: transcribing video files to text with support for multiple model types.
字幕一站式视频处理工具:字幕转换 / 轨道提取封装(MKVToolNix)/ AI 翻译(术语表·断点续传·润色)/ Whisper 转写(双后端)/ 烧录压制,Flutter Windows 桌面版
AI-powered audio transcription with post-processing, supports OpenAI Whisper/GPT and Google Gemini with automatic fallback.
Convert audio and video to text directly on your own machine - 100% local, completely offline, and fully private.
VoicePad is a terminal-first voice note and transcription tool built with Python and Whisper. It records audio, generates local AI-powered transcriptions, and streamlines voice-based workflows through an ergonomic CLI and developer-friendly automation features.
To associate your repository with the transcription-tool topic, visit your repo's landing page and select "manage topics."