Stars
Moshi is a speech-text foundation model and full-duplex spoken dialogue framework. It uses Mimi, a state-of-the-art streaming neural audio codec.
MooER: Moore-threads Open Omni model for speech-to-speech intERaction. MooER-omni includes a series of end-to-end speech interaction models along with training and inference code, covering but not …
中英文敏感词、语言检测、中外手机/电话归属地/运营商查询、名字推断性别、手机号抽取、身份证抽取、邮箱抽取、中日文人名库、中文缩写库、拆字词典、词汇情感值、停用词、反动词表、暴恐词表、繁简体转换、英文模拟中文发音、汪峰歌词生成器、职业名称词库、同义词库、反义词库、否定词库、汽车品牌词库、汽车零件词库、连续英文切割、各种中文词向量、公司名字大全、古诗词库、IT词库、财经词库、成语词库、地名词库、…
BELLE: Be Everyone's Large Language model Engine(开源中文对话大模型)
Google search results crawler, get google search results that you need
Self-Supervised Speech Pre-training and Representation Learning Toolkit
MistSC / s3prl
Forked from s3prl/s3prlSelf-Supervised Speech Pre-training and Representation Learning Toolkit.
🤗 Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and training.
Dynamic Chunk Streaming and Offline Conformer based on athena-team/Athena.
A C++ standalone library for machine learning
HMM, CTC, RNN-Transducer, forward-backward algorithm
A pytorch_lightning reimplementation of the Transducer module from ESPnet.
A curated list of papers dedicated to edit-distance as objective function
Unsupervised text tokenizer for Neural Network-based text generation.
Unsupervised Word Segmentation for Neural Machine Translation and Text Generation
Attempt at tracking states of the arts and recent results (bibliography) on speech recognition.
Python library for Room Impulse Response (RIR) simulation with GPU acceleration
Repo for counting stars and contributing. Press F to pay respect to glorious developers.
It is open source ebook about TensorFlow kernel and implementation mechanism.
DaCiDian is an open-sourced chinese mandarin lexicon for automatic speech recognition(ASR)
lingochamp / kaldi-ctc
Forked from kaldi-asr/kaldiConnectionist Temporal Classification (CTC) Automatic Speech Recognition
Robust Speech Recognition Using Generative Adversarial Networks (GAN)