Skip to content
View huycq1712's full-sized avatar
🚬
focus
🚬
focus
  • Hanoi University of Science and Technology
  • Ha Noi
  • X @huycq1712

Block or report huycq1712

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse

Starred repositories

Showing results

OCR in the Era of Large Language Models

695 50 Updated Aug 12, 2026

A lightweight inference engine supporting speculative speculative decoding (SSD).

Python 986 78 Updated May 10, 2026

A single CLAUDE.md file to improve Claude Code behavior, derived from Andrej Karpathy's observations on LLM coding pitfalls.

201,897 20,715 Updated Apr 20, 2026

This is an open source project (formerly named Listen, Attend and Spell - PyTorch Implementation) for end-to-end ASR implemented with Pytorch, the well known deep learning toolkit.

Python 1,209 314 Updated Dec 19, 2020

Production-grade engineering skills for AI coding agents.

JavaScript 86,606 9,307 Updated Aug 11, 2026
JavaScript 13,906 1,203 Updated May 31, 2026

GLM-OCR: Accurate × Fast × Comprehensive

Python 7,268 658 Updated Apr 21, 2026

Fast and local neural text-to-speech engine

C++ 5,113 485 Updated Aug 12, 2026

A fast, local neural text to speech system

C++ 11,279 1,054 Updated Aug 26, 2025

A fast, local neural text to speech system

C++ 1 Updated Aug 26, 2025

Contexts Optical Compression

Python 23,761 2,194 Updated Jan 27, 2026

A scalable generative AI framework built for researchers and developers working on Large Language Models, Multimodal, and Speech AI (Automatic Speech Recognition and Text-to-Speech)

Python 18,112 3,550 Updated Aug 13, 2026

VietASR - Vietnamese Automatic Speech Recognition

Python 173 59 Updated Jun 18, 2026

[NeurIPS'22] Squeezeformer: An Efficient Transformer for Automatic Speech Recognition

Python 266 19 Updated Feb 12, 2023

Open-Source Toolkit for End-to-End Speech Recognition leveraging PyTorch-Lightning and Hydra.

Python 715 115 Updated Jun 21, 2026

FSA/FST algorithms, differentiable, with PyTorch compatibility.

Cuda 1,350 237 Updated Jul 11, 2026
Python 1,476 425 Updated Jul 16, 2026

A Configurable template for a FastAPI application, with Authentication, User integration, Caching, Logging, Admin pages and a snappy CLI to control it all!

Python 249 17 Updated Aug 7, 2026

Speech-to-text, text-to-speech, speaker diarization, speech enhancement, source separation, and VAD using next-gen Kaldi with onnxruntime without Internet connection. Support embedded systems, Andr…

C++ 14,138 1,624 Updated Aug 11, 2026

Speech-to-text server framework with next-gen Kaldi

C++ 970 146 Updated Aug 13, 2026

SoTA open-source TTS

Python 25,972 3,474 Updated Jul 21, 2026

SGLang is a high-performance serving framework for large language models and multimodal models.

Python 31,718 7,840 Updated Aug 13, 2026

Automatic Speech Recognition (ASR), Speaker Verification, Speech Synthesis, Text-to-Speech (TTS), Language Modelling, Singing Voice Synthesis (SVS), Voice Conversion (VC)

3,127 515 Updated Oct 19, 2023

Port of OpenAI's Whisper model in C/C++

C++ 52,856 6,060 Updated Aug 7, 2026

A mcp server to allow LLMS gain context about shadcn ui component structure,usage and installation,compaitable with react,svelte 5,vue & React Native

TypeScript 2,931 303 Updated May 16, 2026

TTS Dia finetuning for Vietnamese

Python 128 39 Updated Dec 3, 2025

ViStreamASR - Real-Time Vietnamese Speech Recognition

Python 61 21 Updated Jul 12, 2025

The training program for libfacedetection for face detection and 5-landmark detection.

Python 835 213 Updated Jul 10, 2026

An open source library for face detection in images. The face detection speed can reach 1000FPS.

C++ 12,776 3,023 Updated Jun 28, 2026
Next