Skip to content
View ensky0's full-sized avatar

Block or report ensky0

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results
Python 30 3 Updated Jul 21, 2026

This is a Phi Family of SLMs book for getting started with Phi Models. Phi a family of open sourced AI models developed by Microsoft. Phi models are the most capable and cost-effective small langua…

Jupyter Notebook 3,790 512 Updated Aug 11, 2026

Updated list of public BitTorrent trackers

54,882 6,594 Updated Aug 15, 2026

This collection of helper scripts/ and guides for AWS SageMaker HyperPod and ParallelCluster makes it easy to get started with large-scale distributed training on Slurm-based HPC clusters and Kuber…

Python 16 3 Updated Jun 18, 2026

This is a plugin which lets EC2 developers use libfabric as network provider while running NCCL applications.

C++ 231 104 Updated Aug 15, 2026

Documenting my attempts to get the internal speakers of the Galaxy Book4 Pro working on Linux

Shell 21 5 Updated May 20, 2024

Notes and utilities for running Linux on the Samsung Galaxy Book2 Pro

ASL 192 15 Updated Jun 12, 2025

An evolving, large-scale and multi-domain ASR corpus for low-resource languages with automated crawling, transcription and refinement

Python 199 13 Updated Apr 28, 2026

xdotool-like for KDE Plasma

Rust 383 32 Updated Jul 20, 2026

Open Source framework for voice agents, multimodal apps, and realtime AI. Maintained by Daily and the community.

Python 14,140 2,464 Updated Aug 15, 2026

Speech-to-text, text-to-speech, speaker diarization, speech enhancement, source separation, and VAD using next-gen Kaldi with onnxruntime without Internet connection. Support embedded systems, Andr…

C++ 14,200 1,629 Updated Aug 13, 2026

Recurrent neural network for audio noise reduction

C 5,778 1,069 Updated Feb 22, 2025

simple shortcut manager for macOS

Swift 421 16 Updated Jun 3, 2026

CoVoST: A Large-Scale Multilingual Speech-To-Text Translation Corpus (CC0 Licensed)

Python 401 46 Updated Sep 14, 2021

Silero VAD: pre-trained enterprise-grade Voice Activity Detector

Python 9,957 825 Updated Jul 16, 2026

Tools for handling multimodal data in machine learning projects.

Python 1,147 276 Updated Jul 31, 2026

Reference BLEU implementation that auto-downloads test sets and reports a version string to facilitate cross-lab comparisons

Python 1,257 176 Updated Jul 17, 2026

Object-oriented handling of audio data, with GPU-powered augmentations, and more.

Python 351 80 Updated Apr 1, 2025

AcademiCodec: An Open Source Audio Codec Model for Academic Research

Python 674 84 Updated Dec 27, 2023

State-of-the-art audio codec with 90x compression factor. Supports 44.1kHz, 24kHz, and 16kHz mono/stereo audio.

Python 1,842 187 Updated Jul 16, 2026

Foundational Models for State-of-the-Art Speech and Text Translation

Jupyter Notebook 11,839 1,175 Updated Jul 28, 2026

This repository provides a multi-mode and multi-speaker expressive speech synthesis framework, including multi-attentive Tacotron, DurIAN, Non-attentive Tacotron, GST, VAE, GMVAE, and X-vectors for…

Python 74 13 Updated Sep 21, 2022

VocGAN: A High-Fidelity Real-time Vocoder with a Hierarchically-nested Adversarial Network

Python 322 59 Updated Jul 25, 2024

Password cracker for SQLCipher v2 using OpenCL

C 127 34 Updated Jun 5, 2020

SQLCipher is a standalone fork of SQLite that adds 256 bit AES encryption of database files and other security features.

C 7,240 1,398 Updated Jul 8, 2026
Python 46 15 Updated Aug 6, 2025

This is Pytorch Implementation of Google's Non-attentive Tacotron.

Jupyter Notebook 57 13 Updated Dec 21, 2022

PyTorch Implementation of Google's Parallel Tacotron 2: A Non-Autoregressive Neural TTS Model with Differentiable Duration Modeling

Python 191 44 Updated Nov 18, 2021

A statistical model-based Voice Activity Detection

Jupyter Notebook 196 38 Updated Nov 30, 2018
Next