Skip to content
View huycq1712's full-sized avatar
🚬
focus
🚬
focus
  • Hanoi University of Science and Technology
  • Ha Noi
  • X @huycq1712

Block or report huycq1712

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse

Starred repositories

Showing results

OCR in the Era of Large Language Models

694 50 Updated Aug 10, 2026

A lightweight inference engine supporting speculative speculative decoding (SSD).

Python 981 78 Updated May 10, 2026

A single CLAUDE.md file to improve Claude Code behavior, derived from Andrej Karpathy's observations on LLM coding pitfalls.

201,185 20,670 Updated Apr 20, 2026

This is an open source project (formerly named Listen, Attend and Spell - PyTorch Implementation) for end-to-end ASR implemented with Pytorch, the well known deep learning toolkit.

Python 1,209 314 Updated Dec 19, 2020

Production-grade engineering skills for AI coding agents.

JavaScript 85,665 9,224 Updated Aug 8, 2026
JavaScript 13,905 1,204 Updated May 31, 2026

GLM-OCR: Accurate × Fast × Comprehensive

Python 7,261 655 Updated Apr 21, 2026

Fast and local neural text-to-speech engine

C++ 5,090 485 Updated Aug 9, 2026

A fast, local neural text to speech system

C++ 11,277 1,051 Updated Aug 26, 2025

A fast, local neural text to speech system

C++ 1 Updated Aug 26, 2025

Contexts Optical Compression

Python 23,754 2,194 Updated Jan 27, 2026

A scalable generative AI framework built for researchers and developers working on Large Language Models, Multimodal, and Speech AI (Automatic Speech Recognition and Text-to-Speech)

Python 18,082 3,542 Updated Aug 10, 2026

VietASR - Vietnamese Automatic Speech Recognition

Python 173 59 Updated Jun 18, 2026

[NeurIPS'22] Squeezeformer: An Efficient Transformer for Automatic Speech Recognition

Python 266 19 Updated Feb 12, 2023

Open-Source Toolkit for End-to-End Speech Recognition leveraging PyTorch-Lightning and Hydra.

Python 715 115 Updated Jun 21, 2026

FSA/FST algorithms, differentiable, with PyTorch compatibility.

Cuda 1,349 237 Updated Jul 11, 2026
Python 1,473 424 Updated Jul 16, 2026

A Configurable template for a FastAPI application, with Authentication, User integration, Caching, Logging, Admin pages and a snappy CLI to control it all!

Python 249 17 Updated Aug 7, 2026

Speech-to-text, text-to-speech, speaker diarization, speech enhancement, source separation, and VAD using next-gen Kaldi with onnxruntime without Internet connection. Support embedded systems, Andr…

C++ 14,085 1,621 Updated Aug 10, 2026

Speech-to-text server framework with next-gen Kaldi

C++ 971 146 Updated Aug 10, 2026

SoTA open-source TTS

Python 25,947 3,473 Updated Jul 21, 2026

SGLang is a high-performance serving framework for large language models and multimodal models.

Python 31,630 7,781 Updated Aug 10, 2026

Automatic Speech Recognition (ASR), Speaker Verification, Speech Synthesis, Text-to-Speech (TTS), Language Modelling, Singing Voice Synthesis (SVS), Voice Conversion (VC)

3,127 515 Updated Oct 19, 2023

Port of OpenAI's Whisper model in C/C++

C++ 52,792 6,053 Updated Aug 7, 2026

A mcp server to allow LLMS gain context about shadcn ui component structure,usage and installation,compaitable with react,svelte 5,vue & React Native

TypeScript 2,927 302 Updated May 16, 2026

TTS Dia finetuning for Vietnamese

Python 128 39 Updated Dec 3, 2025

ViStreamASR - Real-Time Vietnamese Speech Recognition

Python 61 21 Updated Jul 12, 2025

The training program for libfacedetection for face detection and 5-landmark detection.

Python 835 213 Updated Jul 10, 2026

An open source library for face detection in images. The face detection speed can reach 1000FPS.

C++ 12,773 3,023 Updated Jun 28, 2026
Next