Stars
Official implementation of "MIMO: Controllable Character Video Synthesis with Spatial Decomposed Modeling"
Energy optimization in smart buildings using mixed-integer linear programming using GAMS software
The prediction of average rainfall and average temperature for Iran using machine learning based on historical data
Development of a linear programming code based on the simplex method revised in Pythion software
[CVPR 2024 Highlight] FoundationPose: Unified 6D Pose Estimation and Tracking of Novel Objects
[ECCV 2024] Ray Denoising (RayDN): Depth-aware Hard Negative Sampling for Multi-view 3D Object Detection
Replacing Mamba with xLSTM! It works better. We show that xLSTM-Unet can be an effective semantic segmentation backbone.
[NeurIPS 2024] Code release for "Segment Anything without Supervision"
The official Pytorch Implementation for ElasticDiffusion: Training-free Arbitrary Size Image Generation through Global-Local Content Separation (CVPR 2024)
PixArt-α: Fast Training of Diffusion Transformer for Photorealistic Text-to-Image Synthesis
TheaterGen: Character Management with LLM for Consistent Multi-turn Image Generation
[ECCV 2024] Official PyTorch implementation of "Getting it Right: Improving Spatial Consistency in Text-to-Image Models"
👗 DM-VTON: Distilled Mobile Real-time Virtual Try-On
This is the official repository for the paper "Texture-Preserving Diffusion Models for High-Fidelity Virtual Try-On". CVPR 2024
High-Quality Human Motion Video Generation with Confidence-aware Pose Guidance
SEED-Story: Multimodal Long Story Generation with Large Language Model
[AAAI 2025]👔IMAGDressing👔: Interactive Modular Apparel Generation for Virtual Dressing. It enables customizable human image generation with flexible garment, pose, and scene control, ensuring high …
Real time interactive streaming digital human
Official implementation of EMOPortraits: Emotion-enhanced Multimodal One-shot Head Avatars
[AAAI 2025] EchoMimic: Lifelike Audio-Driven Portrait Animations through Editable Landmark Conditioning
Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation
FaceChain is a deep-learning toolchain for generating your Digital-Twin.