Stars
ForestLee / AudioRecorder
Forked from Dimowner/AudioRecorderAudio Recording Android application
speech-aligner,是一个从“人声语音”及其“语言文本”,产生音素级别时间对齐标注的工具。speech-aligner, is a tool that generate phoneme-level alignment between human speech and its transcription
Silero VAD: pre-trained enterprise-grade Voice Activity Detector
Offline speech recognition API for Android, iOS, Raspberry Pi and servers with Python, Java, C# and Node
stm32f103 RET6 + Marvell88w8801 SDIO Wi-Fi
Microchip Curiosity PIC32MZ - FreeRTOS - LWIP - MBEDTLS
Detecting emotions using MFCC features of human speech using Deep Learning
Open Source Neural Machine Translation and (Large) Language Models in PyTorch
😝 TensorFlowTTS: Real-Time State-of-the-art Speech Synthesis for Tensorflow 2 (supported including English, French, Korean, Chinese, German and Easy to adapt for other languages)
Android Chinese TensorflowTTS demo
kaldi-asr/kaldi is the official location of the Kaldi project.
Ready-to-use OCR with 80+ supported languages and all popular writing scripts including Latin, Chinese, Arabic, Devanagari, Cyrillic and etc.
A Keras implementation of YOLOv3 (Tensorflow backend)
深度学习实践:使用yolov3_keras模型进行实时目标检测(基于Penn-Fudan Database行人数据集
Awesome Object Detection based on handong1587 github: https://handong1587.github.io/deep_learning/2015/10/09/object-detection.html
yolo做行人检测+deep-sort做匹配,端对端做多目标跟踪
多标签文本分类,多标签分类,文本分类, multi-label, classifier, text classification, BERT, seq2seq,attention, multi-label-classification
Convolutional Neural Network for Text Classification in Tensorflow
End-to-end ASR/LM implementation with PyTorch
一个执着于让CPU\端侧-Model逼近GPU-Model性能的项目,CPU上的实时率(RTF)小于0.1
A Deep-Learning-Based Chinese Speech Recognition System 基于深度学习的中文语音识别系统