Lists (1)
Sort Name ascending (A-Z)
Stars
RAG on Paul Graham's essays.
Go from raw audio files to a text-audio dataset automatically with OpenAI's Whisper.
Generate SDKs from Unreal Engine games (UE1 - 4 supported).
Offline speech recognition API for Android, iOS, Raspberry Pi and servers with Python, Java, C# and Node
mp4grep is a CLI for transcribing and searching audio/video files
High-quality implementations of standard and SOTA methods on a variety of tasks.
ROS node for speech to text using VOSK on a docker container
extract text from any document. no muss. no fuss.
Code for the ICASSP-2021 paper: Continuous Speech Separation with Conformer.
OpenVINO™ is an open source toolkit for optimizing and deploying AI inference
Avalanche: an End-to-End Library for Continual Learning based on PyTorch.
Label Studio is a multi-type data labeling and annotation tool with standardized output format
The SpeechBrain project aims to build a novel speech toolkit fully based on PyTorch. With SpeechBrain users can easily create speech processing systems, ranging from speech recognition (both HMM/DN…
Persists tmux environment across system restarts.
PyTorch implementation of LF-MMI for End-to-end ASR
Offline extractor of synchronous context-free grammars for machine translation.
Tools to download and cleanup Common Crawl data
MATLAB real-time/interactive speech tools. This series is obsolete. SP3ARK is the up-to-date series (will be).
Open-Unmix - Music Source Separation for PyTorch
An open-source speech separation and enhancement library
Evaluate results from ASR/Speech-to-Text quickly
Github Pages template based upon HTML and Markdown for personal, portfolio-based websites.
[ICASSP20] A Dialogical Emotion Decoder For Speech Emotion Recognition in Spoken Dialog