Skip to content
View Chuny1's full-sized avatar

Block or report Chuny1

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse

Starred repositories

Showing results

[CVPR 2026 Oral] "MARCO: Navigating the Unseen Space of Semantic Correspondence"

Python 148 6 Updated Apr 21, 2026

A data collection and processing pipeline for animal video, annotations include mask, keypoint, depth, occlusion, etc. Suitable for 3D/4D reconstruction, tracking, pose prediction, etc.

Python 63 3 Updated Dec 5, 2025

[ICLR 2026 oral] Official code for VIST3A: Text-to-3D by Stitching a Multi-view Reconstruction Network to a Video Generator

Python 144 4 Updated May 20, 2026

🏂 Training-Free Human Mesh Recovery from Videos, based on SAM-3, Diffusion-VAS, and SAM-3D-Body.

Python 371 37 Updated May 11, 2026

Native and Compact Structured Latents for 3D Generation

Python 10,675 1,288 Updated Jul 10, 2026

Official implementation of the 2024 ECCV paper SHIC: Shape-Image Correspondences with no Keypoint Annotation

Jupyter Notebook 39 2 Updated Oct 1, 2024

[ECCV 2024] Official implementation of the paper "X-Pose: Detecting Any Keypoints"

Python 818 45 Updated Aug 16, 2024

[CVPR 2026 Findings] TalkVid: A Large-Scale Diversified Dataset for Audio-Driven Talking Head Synthesis

Python 197 13 Updated Jun 8, 2026

Muti-human Interactive Talking Dataset

Python 76 1 Updated Aug 6, 2025

Official implementation of "MoMask: Generative Masked Modeling of 3D Human Motions (CVPR2024)"

Python 1,303 115 Updated Sep 13, 2024

The ultimate training toolkit for finetuning diffusion models

Python 11,747 1,492 Updated Aug 17, 2026

Controllable video and image Generation, SVD, Animate Anyone, ControlNet, ControlNeXt, LoRA

Python 1,646 80 Updated Sep 25, 2024

Code and models for ICML 2024 paper, NExT-GPT: Any-to-Any Multimodal Large Language Model

Python 3,637 359 Updated May 13, 2025

A PyTorch library for implementing flow matching algorithms, featuring continuous and discrete flow matching implementations. It includes practical examples for both text and image modalities.

Python 4,688 369 Updated Jan 5, 2026

Unified Efficient Fine-Tuning of 100+ LLMs & VLMs (ACL 2024)

Python 74,185 9,077 Updated Aug 13, 2026

[CVPR 2024] "LL3DA: Visual Interactive Instruction Tuning for Omni-3D Understanding, Reasoning, and Planning"; an interactive Large Language 3D Assistant.

Python 320 14 Updated Jul 17, 2024

Official Implementation of Diffusion Step Annealing (DiSA) in Autoregressive Image Generation

Jupyter Notebook 143 1 Updated May 27, 2025

This repository contains an implementation for performing 3D animal (quadruped) reconstruction from a monocular image or video. The system adapts the pose (limb positions) and shape (animal type/he…

Python 153 23 Updated Nov 2, 2024

[IJCV 2022] Bridging Composite and Real: Towards End-to-end Deep Image Matting

Python 940 134 Updated Apr 13, 2023

Official implementation of DeepLabCut: Markerless pose estimation of user-defined features with deep learning for all animals incl. humans

Python 5,736 1,790 Updated Aug 17, 2026

official implementation of paper "Tactile DreamFusion: Exploiting Tactile Sensing for 3D Generation"

Jupyter Notebook 57 6 Updated Apr 15, 2026

2027 AI/ML internship & new graduate job list updated daily

6,142 239 Updated Aug 17, 2026

🚀 A simple way to launch, train, and use PyTorch models on almost any device and distributed configuration, automatic mixed precision (including fp8), and easy-to-configure FSDP and DeepSpeed support

Python 9,820 1,435 Updated Aug 17, 2026

[CVPR 2024 - Oral, Best Paper Award Candidate] Marigold: Repurposing Diffusion-Based Image Generators for Monocular Depth Estimation

Python 3,195 208 Updated Dec 10, 2025

[CVPR2024 (Highlight)] RichDreamer: A Generalizable Normal-Depth Diffusion Model for Detail Richness in Text-to-3D. Live Demo:https://modelscope.cn/studios/Damo_XR_Lab/3D_AIGC

Python 477 23 Updated Sep 27, 2024

naive filter of objaverse

Python 152 2 Updated Mar 15, 2024

Let us control diffusion models!

Python 34,073 3,018 Updated Feb 25, 2024

[3DV 2025 Best Paper] We present Object Images (Omages): An homage to the classic Geometry Images.

Jupyter Notebook 370 15 Updated Jan 2, 2025

[ECCV 2024] Code for VFusion3D: Learning Scalable 3D Generative Models from Video Diffusion Models

Python 457 36 Updated Sep 9, 2024

[ICLR 2025] HD-Painter: High-Resolution and Prompt-Faithful Text-Guided Image Inpainting with Diffusion Models

Python 368 23 Updated Mar 14, 2024
Next