Skip to content
View tyfeld's full-sized avatar

Highlights

  • Pro

Organizations

@Gen-Verse

Block or report tyfeld

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

Official Repo For PerceptionDLM Codebase

Python 77 4 Updated Jun 22, 2026

[ICML 2026] The official implementation of paper "Unified Multimodal Autoregressive Modeling with Shared Context—Visual Tokenizer is Key to Unification"

Python 46 Updated Jul 13, 2026

ARM: An AutoRegressive Large Multimodal Model with Discrete Representations

50 Updated Jun 10, 2026

Bernini is a unified framework for video generation and editing that combines an MLLM-based semantic planner with a DiT-based renderer.

Python 1,174 88 Updated Jul 13, 2026

A 3B-active-parameter native unified multimodal model for image and video understanding, generation, and editing.

Python 1,291 91 Updated Jul 14, 2026

[ICLR'26] Traceable Evidence Enhanced Visual Grounded Reasoning: Evaluation and Methodology

Python 92 1 Updated Jan 26, 2026

VideoFlexTok: Flexible-Length Coarse-to-Fine Video Tokenization

Jupyter Notebook 43 1 Updated Apr 16, 2026

Official Codebase For paper "One-step Language Modeling via Continuous Denoising"

Python 160 11 Updated Jul 7, 2026

Skill package for ML/CV/NLP paper writing, curated and adapted from Prof. Peng Sida's open notes for Codex, Claude Code, and Gemini.

5,463 269 Updated Jun 23, 2026

[🚀 ICLR 2026 Oral] NextStep-1: SOTA Autogressive Image Generation with Continuous Tokens. A research project developed by the StepFun’s Multimodal Intelligence team.

Python 689 26 Updated Feb 27, 2026

Personal PyTorch implementation of "Generative Modeling via Drifting" with Claude

Python 240 24 Updated Feb 6, 2026

Elevate your AI research writing, no more tedious polishing ✨

31,997 2,400 Updated May 18, 2026

Official repo for UAE

Python 207 8 Updated Jun 21, 2026

dLLM: Simple Diffusion Language Modeling

Python 2,651 280 Updated Jul 17, 2026

A paper list of Awesome Latent Space.

947 39 Updated Jul 13, 2026

Official code implementation of Context Cascade Compression: Exploring the Upper Limits of Text Compression

Python 313 6 Updated Jan 27, 2026

PyTorch implementation of JiT https://arxiv.org/abs/2511.13720

Python 2,465 163 Updated Dec 8, 2025

Official Implementation of "MMaDA-Parallel: Multimodal Large Diffusion Language Models for Thinking-Aware Editing and Generation"

Python 301 10 Updated Jan 29, 2026
Jupyter Notebook 145 8 Updated Nov 8, 2025

[ICLR 2026 Oral & ICML 2026] Generative Universal Verifier as Multimodal Meta-Reasoner

Python 64 7 Updated May 29, 2026

Official implementation of "Continuous Autoregressive Language Models"

Python 811 92 Updated May 7, 2026

GPU-optimized framework for training diffusion language models at any scale. The backend of Quokka, Super Data Learners, and OpenMoE 2 training.

Python 343 35 Updated Nov 11, 2025

A Curated List of Awesome Works in World Modeling, Aiming to Serve as a One-stop Resource for Researchers, Practitioners, and Enthusiasts Interested in World Modeling.

3,212 136 Updated Jul 20, 2026

[NeurIPS 2025] Encoder-Decoder Diffusion Language Models for Efficient Training and Inference

Python 47 3 Updated Oct 29, 2025

The official implementation of dLLM-Var

Python 35 1 Updated Nov 6, 2025

[ICLR 2026] 🐻 Uniform Discrete Diffusion with Metric Path for Video Generation

Python 123 5 Updated May 20, 2026

[ICLR'26] Grasp Any Region: Towards Precise, Contextual Pixel Understanding for Multimodal LLMs

Python 99 10 Updated Jan 26, 2026
Next