Skip to content
View wanglamao's full-sized avatar

Block or report wanglamao

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

SOTA-Class TTS at Compact Scale

Python 607 50 Updated Aug 9, 2026

An MCP server that lets LLM agents play Civilization VI.

Python 140 26 Updated Jun 21, 2026

VulnGym: A Real-World, Project-Level Vulnerability Benchmark for White-Box Vulnerability-Hunting Agents

231 41 Updated Jun 26, 2026

Code for paper "The Philosopher’s Stone: Trojaning Plugins of Large Language Models"

Python 33 5 Updated Sep 11, 2024

Industrial audio online policy distillation (OPD) training stack for ASR and TTS, distilling compact audio models from stronger teacher models.

Python 909 58 Updated Jun 5, 2026

A real-time and multilingual speech translation model

Python 268 23 Updated Feb 13, 2026

Video stabilization using gyroscope data

Rust 9,330 482 Updated Jul 16, 2026

🏭 Mega Scale Multimodal DataPipeline for SOTA Foundation Models

Python 375 46 Updated May 12, 2026

Kyutai's Speech-To-Text and Text-To-Speech models based on the Delayed Streams Modeling framework.

Python 3,014 313 Updated Jan 26, 2026

Research of DeepSeek Engram Architecture based on Qwen-3 and Stable Diffusion series.

Python 1,083 73 Updated May 21, 2026

vLLM Adaptation for GLM ASR Nano

Python 3 Updated Jan 22, 2026

[AutoArk] GPA (General Purpose Audio) can do ASR, TTS and voice conversion with one tiny model!

Python 2,468 167 Updated May 25, 2026

EVA OS — A real-time multimodal AIOS for next-generation hardware, enabling your devices being “alive” and as intelligent as a real brain.

1,565 103 Updated Aug 3, 2026

Reference PyTorch implementation and models for DINOv3

Jupyter Notebook 11,202 932 Updated Jul 15, 2026

Latent Consistency Models: Synthesizing High-Resolution Images with Few-Step Inference

Python 4,619 232 Updated Jun 14, 2024

Multi-lingual large voice generation model, providing inference, training and deployment full-stack ability.

Python 22,818 2,629 Updated May 25, 2026

Fast inference engine for Transformer models

C++ 4,626 515 Updated Aug 16, 2026

Build userspace NVMe drivers and storage applications with CUDA support

C 444 56 Updated Dec 18, 2023

[CVPR 2024] This is the official source for our paper "SyncTalk: The Devil is in the Synchronization for Talking Head Synthesis"

Python 1,626 196 Updated Sep 18, 2025

GaLore: Memory-Efficient LLM Training by Gradient Low-Rank Projection

Python 1,700 167 Updated Oct 28, 2024

RepViT: Revisiting Mobile CNN From ViT Perspective [CVPR 2024] and RepViT-SAM: Towards Real-Time Segmenting Anything

Jupyter Notebook 1,108 83 Updated Jun 14, 2024

[CVPR 2023] L2G-NeRF: Local-to-Global Registration for Bundle-Adjusting Neural Radiance Fields

Python 253 6 Updated Mar 6, 2024

Infinite Photorealistic Worlds using Procedural Generation

Python 7,209 608 Updated Aug 5, 2026

IPTV直播源抓取 自动整合hao趣网直播源+TVBox直播源+其他网上直播源 择取分辨率、速度最佳视频流 定期更新

10,279 698 Updated Dec 31, 2024

Digital Avatar Conversational System - Linly-Talker. 😄✨ Linly-Talker is an intelligent AI system that combines large language models (LLMs) with visual models to create a novel human-AI interaction…

Python 3,434 534 Updated Feb 10, 2026

Unity project for nerf_pl (Neural Radiance Fields)

C# 221 21 Updated Nov 3, 2021

InstantID: Zero-shot Identity-Preserving Generation in Seconds 🔥

Python 11,987 885 Updated Jul 18, 2024

Code release for NeRF (Neural Radiance Fields)

Jupyter Notebook 10,929 1,431 Updated Apr 12, 2025

An interface for text-guided mesh refinement.

Python 188 14 Updated Jan 27, 2024
Next