Skip to content
View dhgrs's full-sized avatar
  • Japan

Organizations

@monthly-hack

Block or report dhgrs

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

An interface library for RL post training with environments.

Python 2,598 455 Updated Sep 19, 2026

Qwen3-TTS is an open-source series of TTS models developed by the Qwen team at Alibaba Cloud, supporting stable, expressive, and streaming speech generation, free-form voice design, and vivid voice…

Python 13,476 1,743 Updated Mar 17, 2026

JATTS: A modern, research-oriented Japanese Text-to-speech Open-sourced Toolkit

Python 44 1 Updated Mar 13, 2026

A Datacenter Scale Distributed Inference Serving Framework

Rust 8,135 1,612 Updated Sep 21, 2026

Production-tested AI infrastructure tools for efficient AGI development and community-driven innovation

8,066 295 Updated May 15, 2025

Achieve the llama3 inference step-by-step, grasp the core concepts, master the process derivation, implement the code.

Jupyter Notebook 629 53 Updated Feb 24, 2025

A high-throughput and memory-efficient inference and serving engine for LLMs

Python 92,274 22,458 Updated Sep 21, 2026

SGLang is a high-performance serving framework for large language models and multimodal models.

Python 36,230 9,024 Updated Sep 21, 2026

TensorRT LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and supports state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs. Tensor…

Python 14,682 2,765 Updated Sep 21, 2026

pyopenjtalk-plus: A Python wrapper for OpenJTalk with additional improvements

Python 64 5 Updated Aug 11, 2026

[ICASSP 2024] TextrolSpeech: A Text Style Control Speech Corpus With Codec Language Text-to-Speech Models

Python 188 5 Updated Nov 22, 2024

Large-scale Self-supervised Pre-training Across Tasks, Languages, and Modalities

Python 22,219 2,706 Updated Sep 21, 2026

Official implementation of the paper "BigCodec: Pushing the Limits of Low-Bitrate Neural Speech Codec"

Python 221 20 Updated Sep 19, 2024

Maestro: Netflix’s Workflow Orchestrator

Java 3,839 310 Updated Sep 16, 2026

Controllable and fast Text-to-Speech for over 7000 languages!

Python 2,209 317 Updated Jan 25, 2026
Python 401 68 Updated Sep 3, 2024

Inference and training library for high-quality TTS models.

Python 5,593 585 Updated Dec 10, 2024

Implementation of Band Split Roformer, SOTA Attention network for music source separation out of ByteDance AI Labs

Python 949 46 Updated Jun 14, 2026

A PyTorch-based Speech Toolkit

Python 11,829 1,725 Updated Aug 27, 2026

The Nendo AI Audio Tool Suite

Python 218 11 Updated Apr 25, 2024

Nendo is an open source platform for AI-driven audio management, intelligence, and generation.

Makefile 129 11 Updated Mar 20, 2024

Starter-kit to build constrained agents with Nextjs, FastAPI and Langchain

TypeScript 1,948 270 Updated Mar 18, 2026

🔊 Text-Prompted Generative Audio Model

Jupyter Notebook 39,272 4,669 Updated Aug 19, 2024

青空文庫振り仮名注釈付き音声コーパスのデータセット

50 1 Updated Mar 7, 2025

Foundational model for human-like, expressive TTS

Python 4,207 689 Updated Jul 30, 2024

[ICASSP 2024] This is the official code for "VoiceFlow: Efficient Text-to-Speech with Rectified Flow Matching"

Python 374 23 Updated Sep 3, 2024

Easily use and train state of the art late-interaction retrieval methods (ColBERT) in any RAG pipeline. Designed for modularity and ease-of-use, backed by research.

Python 3,958 276 Updated May 17, 2025
Next