Skip to content
View xwen99's full-sized avatar

Highlights

  • Pro

Organizations

@CVMI-Lab

Block or report xwen99

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

Official Implemenation for RAEv2: Improved Baselines with Representation Autoencoders

Python 315 13 Updated May 21, 2026

Cambrian-P: Pose-Grounded Video Understanding

Python 104 3 Updated Jul 28, 2026

Claude code environment for laywers

Shell 734 71 Updated Apr 25, 2026
Python 950 75 Updated Jun 26, 2026

Accompanying code for "Discovering State-of-the-art Reinforcement Algorithms" Nature publication

Python 720 60 Updated Dec 2, 2025

GameWorld: Towards Standardized and Verifiable Evaluation of Multimodal Game Agents

Python 215 9 Updated Jul 31, 2026

JoyAI-Image is the unified multimodal foundation model for image understanding, text-to-image generation, and instruction-guided image editing.

Python 2,255 160 Updated Aug 5, 2026
Python 94 1 Updated May 8, 2026

Project website for "Beyond Language Modeling: An Exploration of Multimodal Pretraining" paper.

HTML 9 Updated Mar 4, 2026

The first multiplayer video world model in Minecraft

Python 221 9 Updated Mar 3, 2026

Sample LaTex file for HKU PhD thesis.

TeX 30 4 Updated Mar 16, 2022

LaTeX Template for HKU MPhil and PhD Thesis

TeX 48 18 Updated Feb 11, 2019

Scaling Text-to-Image Diffusion Transformers with Representation Autoencoders

Python 255 4 Updated Feb 13, 2026

[CVPR 2026] Pixio: a capable vision encoder dedicated to dense prediction, simply by pixel reconstruction

Python 475 12 Updated Jul 17, 2026

Official code of "LayerTracer: Cognitive-Aligned Layered SVG Synthesis via Diffusion Transformer"

Python 98 3 Updated Apr 1, 2025

Data set of Sudoku puzzles

102 12 Updated Sep 22, 2023

Sudoku4LLM is a Sudoku dataset generator for training and evaluating reasoning in Large Language Models (LLMs). It offers customizable puzzles, difficulty levels, and 11 serialization formats to su…

Python 10 Updated Apr 28, 2025

PyTorch implementation of JiT https://arxiv.org/abs/2511.13720

Python 2,489 165 Updated Dec 8, 2025

Official implementation of "Continuous Autoregressive Language Models"

Python 814 92 Updated May 7, 2026

EB1A Green Card Template for Self-Petition

TeX 93 39 Updated Jun 16, 2026

Instella-T2I: Pushing the Limits of 1D Discrete Latent Space Image Generation

Python 25 Updated May 14, 2026

The repository provides code for running inference and finetuning with the Meta Segment Anything Model 3 (SAM 3), links for downloading the trained model checkpoints, and example notebooks that sho…

Python 11,260 1,700 Updated Jul 31, 2026

A Walsh Hadamard Derived Linear Vector Symbolic Architecture 🔥

Python 12 3 Updated Jan 30, 2026

[NeurIPS'25] Official repository of Concerto: Joint 2D-3D Self-Supervised Learning Emerges Spatial Representations

Python 533 29 Updated Apr 7, 2026

Contexts Optical Compression

Python 23,758 2,194 Updated Jan 27, 2026

Official PyTorch Implementation of "Diffusion Transformers with Representation Autoencoders"

Python 1,988 87 Updated Feb 25, 2026

[ICLR 2026]QeRL enables RL for 32B LLMs on a single H100 GPU.

Python 512 52 Updated Mar 30, 2026

A comprehensive JAX/NNX library for diffusion and flow matching generative algorithms, featuring DiT (Diffusion Transformer) and its variants as the primary backbone with support for ImageNet train…

Python 153 13 Updated Oct 16, 2025

(NeurIPS 2025) Vision Foundation Models as Effective Visual Tokenizers for Autoregressive Image Generation

Python 77 Updated May 21, 2026
Next