-
Bytedance (Tiktok)
- Singapore
- https://lxtgh.github.io/
- @xtl994
Highlights
- Pro
Lists (3)
Sort Name ascending (A-Z)
Stars
《深入理解 AI Agent:设计原理与工程实践》(李博杰 著)开源主仓库:全书正文、编译版 PDF 与按章配套代码
EdgeBench: Unveiling scaling laws of learning from real-world environments
A Curated List of Vision-Language-Action (VLA) and World Action Models (WAM) Research and Beyond
SANA: Efficient High-Resolution Image Synthesis with Linear Diffusion Transformer
JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents
Domain-specific language designed to streamline the development of high-performance GPU/CPU/Accelerators kernels
Unified World Model Inference & Evaluation Infrastructure
🌐 Vision-Language-Action Models for Autonomous Driving: Past, Present, and Future
Official Python inference and LoRA trainer package for the LTX-2 audio–video generative model.
Matrix-Game 3.5: Enhancing Real-Time Streaming Interactive World Models with Patch Memory
Offical repo for the paper "LATO.2: Factorized 3D Mesh Generation with Vertex and Topology Flow"
The end of web parsing. The beginning of scalable pixel-native search. link: https://pixelrag.ai/
Sutskever 30 implementations inspired by https://papercode.vercel.app/ | For Agents, use https://github.com/pageman/Sutskever-Agent | Polyglot / Multi-Backed version at https://github.com/pageman/s…
Official Repo For PerceptionDLM Codebase
The open-source CapCut alternative
[ECCV 2026] Official code for paper "Actor as Its Own Critic: Unifying Region Understanding and Localization via CycleGRPO"
Vision as Unified Multimodal Generation
The Source Code for OmniVideoBench @ICLR 2026
[ECCV 2026] Official code for paper: MotionAtlas: Detailed Region Captioning for Motion-Centric Videos
Flexible and Pluggable Serving Engine for Diffusion LLMs
Qwen-AgentWorld: Language World Models for General Agents
🚀 Pytorch Distributed native training library for LLMs/VLMs with OOTB Hugging Face support