Skip to content
View Aleafy's full-sized avatar
🎯
Focusing
🎯
Focusing

Block or report Aleafy

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

HY-SOAR:Self-Correction for Optimal Alignment and Refinement in Diffusion Models

Python 741 64 Updated Apr 21, 2026

A 3B-active-parameter native unified multimodal model for image and video understanding, generation, and editing.

Python 1,320 94 Updated Jul 14, 2026

Project Lyra: Open Generative 3D World Models

Python 2,240 227 Updated Jul 20, 2026

A unified inference and post-training framework for accelerated video generation.

Python 3,943 402 Updated Aug 14, 2026

Unified Codebase for Advanced World Models.

Python 853 50 Updated Aug 3, 2026

An in-the-wild benchmark for AI agents in the OpenClaw Environment.

Python 510 58 Updated Aug 13, 2026

EasyR1: An Efficient, Scalable, Multi-Modality RL Training Framework based on veRL

Python 5,115 383 Updated Jul 30, 2026

ARIS ⚔️ (Auto-Research-In-Sleep) — Lightweight Markdown-only skills for autonomous ML research: cross-model review loops, idea discovery, and experiment automation. No framework, no lock-in — works…

Python 14,685 1,294 Updated Aug 11, 2026

🤗 LeRobot: Making AI for Robotics more accessible with end-to-end learning

Python 26,654 5,385 Updated Aug 14, 2026

VIGA: Vision-as-Inverse-Graphics Agent

Python 1,270 124 Updated May 6, 2026

A unified framework for easy reinforcement learning in Flow-Matching models

Python 658 53 Updated Aug 13, 2026

This repository provides FlashPortrait custom nodes for ComfyUI.

Python 25 2 Updated Dec 29, 2025

A Unified Visual Generator with Interleaved OmniModal Context

Python 233 3 Updated Mar 5, 2026

[ICLR 26 Oral] Stable Video Infinity: Infinite-Length Video Generation with Error Recycling

Python 2,549 226 Updated Jun 3, 2026

Official code for StoryMem: Multi-shot Long Video Storytelling with Memory

Python 761 75 Updated Jul 22, 2026

The official implementation of InfiniteVGGT

Python 382 21 Updated Apr 19, 2026

Mixture-of-Groups Attention for End-to-End Long Video Generation

100 Updated Oct 22, 2025

The repository provides code for running inference with the Meta Segment Anything Audio Model (SAM-Audio), links for downloading the trained model checkpoints, and example notebooks that show how t…

Python 3,592 326 Updated May 26, 2026

Qwen-Image-Layered: Layered Decomposition for Inherent Editablity

Python 2,065 167 Updated Dec 31, 2025

Official Implementation of "MemFlow: Flowing Adaptive Memory for Consistent and Efficient Long Video Narratives"

Python 216 9 Updated Dec 29, 2025

[NeurIPS 24] The implementation and dataset of LiveScene: Language Embedding Interactive Radiance Fields for Physical Scene Rendering and Control

Python 60 2 Updated Mar 31, 2025

[Siggraph Asia 25] SS4D: Native 4D Generative Model via Structured Spacetime Latents

Python 36 3 Updated Dec 17, 2025

HY-World 1.5: A Systematic Framework for Interactive World Modeling with Real-Time Latency and Geometric Consistency

Python 1,575 142 Updated Jun 10, 2026

[CVPR2026]We present FlashPortrait, an end-to-end video diffusion transformer capable of synthesizing ID-preserving, infinite-length videos while achieving up to 6$\times$ acceleration in inference…

Python 482 38 Updated Feb 21, 2026

[CVPR 2026] V-RGBX: Video Editing with Accurate Controls over Intrinsic Properties

Python 110 7 Updated Jan 17, 2026

[CVPR 2026] Official Code for "ARM-Thinker: Reinforcing Multimodal Generative Reward Models with Agentic Tool Use and Visual Reasoning"

Python 196 12 Updated Feb 13, 2026

ViSAudio: End-to-End Video-Driven Binaural Spatial Audio Generation

118 4 Updated Dec 11, 2025

[ICLR 2026] An official implementation of "STAR-Bench: Probing Deep Spatio-Temporal Reasoning as Audio 4D Intelligence"

Python 43 4 Updated Apr 19, 2026

CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning

Python 37 3 Updated Aug 28, 2025

We present StableAvatar, the first end-to-end video diffusion transformer, which synthesizes infinite-length high-quality audio-driven avatar videos without any post-processing, conditioned on a re…

Python 1,255 113 Updated Jan 20, 2026
Next