Skip to content
View Atotti's full-sized avatar

Highlights

  • Pro

Organizations

@citruzdev @gdsc-tmu @triC-tmu

Block or report Atotti

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

Custom end4

QML 648 29 Updated Jul 26, 2026

Benchmarking STT service TTFB and semantic WER for real-time AI applications

Python 94 25 Updated Jul 25, 2026

Coco-Nut (Corpus of connecting NIHONGO utterance and text) corpus

21 Updated Jun 12, 2024

Post-training with Tinker

Python 3,915 491 Updated Jul 26, 2026

Powerful system-level package manager for Linux, macOS and Windows written in Rust – building on top of the Conda ecosystem.

Rust 7,506 545 Updated Jul 26, 2026

A fast, cross-platform build tool inspired by Make, designed for modern workflows.

Go 15,879 871 Updated Jul 26, 2026

AIST Toolkit for Accelerating Machine Learning Research

Python 44 4 Updated Jul 26, 2026
Python 40 7 Updated Jul 6, 2026
Jupyter Notebook 11 7 Updated Jul 23, 2026

agent multiplexer that lives in your terminal.

Rust 20,913 1,398 Updated Jul 26, 2026

A toolkit for speaker diarization.

Jupyter Notebook 505 60 Updated May 29, 2026
Python 30 3 Updated Jan 5, 2026

MoshiRAG is a compact full-duplex speech language model augmented with asynchronous knowledge retrieval to improve factuality without sacrificing real-time interactivity.

Rust 142 11 Updated Apr 28, 2026
Python 29 2 Updated Jul 3, 2026
Python 6 Updated Jan 7, 2026

FlashCosyVoice: A lightweight vLLM implementation built from scratch for CosyVoice.

Python 250 25 Updated Feb 25, 2026

The agent that grows with you

Python 220,663 42,023 Updated Jul 26, 2026
Python 97 11 Updated Oct 23, 2024

Large-scale, Informative, and Diverse Multi-round Chat Data (and Models)

Python 2,875 142 Updated Mar 13, 2024

Reference implementation of an end-to-end voice agent built using the NVIDIA Nemotron models

Python 166 46 Updated Jul 24, 2026

[EMNLP 2025 Findings] Code for "Distilling Many-Shot In-Context Learning into a Cheat Sheet"

Python 6 Updated Nov 21, 2025

List of open-source TTS, voice cloning, and music generation models

393 54 Updated Jul 23, 2026
Python 30 8 Updated Jul 16, 2026
Python 105 17 Updated Jul 18, 2026

Erasing concepts from neural representations with provable guarantees

Python 258 15 Updated Jan 27, 2025

Must-read Papers on LLM Agents.

3,089 184 Updated Jul 5, 2026

A package for NeuCodec: a 50hz, 0.8kbps, 24kHz audio codec.

Python 161 27 Updated Jun 22, 2026

MOSS‑TTS Family is an open‑source speech and sound generation model family from MOSI.AI and the OpenMOSS team. It is designed for high‑fidelity, high‑expressiveness, and complex real‑world scenario…

Python 3,902 348 Updated Jun 22, 2026

Whisper-Flow is a framework designed to enable real-time transcription of audio content using OpenAI’s Whisper model. Rather than processing entire files after upload (“batch mode”), Whisper-Flow a…

Python 815 119 Updated Jul 26, 2026
Next