Skip to content
View xwen99's full-sized avatar

Highlights

  • Pro

Organizations

@CVMI-Lab

Block or report xwen99

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

Official Implemenation for RAEv2: Improved Baselines with Representation Autoencoders

Python 310 13 Updated May 21, 2026

Cambrian-P: Pose-Grounded Video Understanding

Python 102 3 Updated Jul 17, 2026

Claude code environment for laywers

Shell 726 69 Updated Apr 25, 2026
Python 940 74 Updated Jun 26, 2026

Accompanying code for "Discovering State-of-the-art Reinforcement Algorithms" Nature publication

Python 717 61 Updated Dec 2, 2025

GameWorld: Towards Standardized and Verifiable Evaluation of Multimodal Game Agents

Python 213 8 Updated Jul 21, 2026

JoyAI-Image is the unified multimodal foundation model for image understanding, text-to-image generation, and instruction-guided image editing.

Python 2,245 158 Updated Jul 17, 2026
Python 92 1 Updated May 8, 2026

Project website for "Beyond Language Modeling: An Exploration of Multimodal Pretraining" paper.

HTML 9 Updated Mar 4, 2026

The first multiplayer video world model in Minecraft

Python 219 9 Updated Mar 3, 2026

Sample LaTex file for HKU PhD thesis.

TeX 30 4 Updated Mar 16, 2022

LaTeX Template for HKU MPhil and PhD Thesis

TeX 48 18 Updated Feb 11, 2019

Scaling Text-to-Image Diffusion Transformers with Representation Autoencoders

Python 255 4 Updated Feb 13, 2026

[CVPR 2026] Pixio: a capable vision encoder dedicated to dense prediction, simply by pixel reconstruction

Python 461 12 Updated Jul 17, 2026

Official code of "LayerTracer: Cognitive-Aligned Layered SVG Synthesis via Diffusion Transformer"

Python 98 3 Updated Apr 1, 2025

Data set of Sudoku puzzles

98 11 Updated Sep 22, 2023

Sudoku4LLM is a Sudoku dataset generator for training and evaluating reasoning in Large Language Models (LLMs). It offers customizable puzzles, difficulty levels, and 11 serialization formats to su…

Python 10 Updated Apr 28, 2025

PyTorch implementation of JiT https://arxiv.org/abs/2511.13720

Python 2,471 163 Updated Dec 8, 2025

Official implementation of "Continuous Autoregressive Language Models"

Python 812 92 Updated May 7, 2026

EB1A Green Card Template for Self-Petition

TeX 92 38 Updated Jun 16, 2026

Instella-T2I: Pushing the Limits of 1D Discrete Latent Space Image Generation

Python 25 Updated May 14, 2026

The repository provides code for running inference and finetuning with the Meta Segment Anything Model 3 (SAM 3), links for downloading the trained model checkpoints, and example notebooks that sho…

Python 11,100 1,675 Updated Jul 15, 2026

A Walsh Hadamard Derived Linear Vector Symbolic Architecture 🔥

Python 12 3 Updated Jan 30, 2026

[NeurIPS'25] Official repository of Concerto: Joint 2D-3D Self-Supervised Learning Emerges Spatial Representations

Python 531 29 Updated Apr 7, 2026

Contexts Optical Compression

Python 23,687 2,186 Updated Jan 27, 2026

Official PyTorch Implementation of "Diffusion Transformers with Representation Autoencoders"

Python 1,978 86 Updated Feb 25, 2026

[ICLR 2026]QeRL enables RL for 32B LLMs on a single H100 GPU.

Python 511 52 Updated Mar 30, 2026

A comprehensive JAX/NNX library for diffusion and flow matching generative algorithms, featuring DiT (Diffusion Transformer) and its variants as the primary backbone with support for ImageNet train…

Python 152 13 Updated Oct 16, 2025

(NeurIPS 2025) Vision Foundation Models as Effective Visual Tokenizers for Autoregressive Image Generation

Python 77 Updated May 21, 2026
Next