Stars
Does VLA Even Know the Basics? Measuring Commonsense and World Knowledge Retention in Vision-Language-Action Models
Web-based augmented reality (AR) tool for animating research posters via phone scanning
CADEvolve: Creating Realistic CAD via Program Evolution
[ACL2026] Z3D: Zero-Shot 3D Visual Grounding from Images
[AAAI 2026] BREPS: Bounding-Box Robustness Evaluation of Promptable Segmentation
[EACL 2026] SPARTA: Evaluating Reasoning Segmentation Robustness through Black-Box Adversarial Paraphrasing in Text Autoencoder Latent Space
[CVPR2026] Zoo3D: Zero-Shot 3D Object Detection at Scene Level
[ICRA2026] TUN3D: Towards Real-World Scene Understanding from Unposed Images
[WACV2022] ImVoxelNet: Image to Voxels Projection for Monocular and Multi-View General-Purpose 3D Object Detection
This is the official implementation of "ImageReFL: Balancing Quality and Diversity in Human-Aligned Diffusion Models"
[ICLR2026] cadrille: Multi-modal CAD Reconstruction with Online Reinforcement Learning
[ICCV2025] CAD-Recode: Reverse Engineering CAD Code from Point Clouds
[NeurIPS 2024] RClicks: Realistic Click Simulation for Benchmarking Interactive Segmentation
[AAAI2025] UniDet3D: Multi-dataset Indoor 3D Object Detection
Official Implementation for "The Devil is in the Details: StyleFeatureEditor for Detail-Rich StyleGAN Inversion and High Quality Image Editing"
[CVPR2024] OneFormer3D: One Transformer for Unified Point Cloud Segmentation
Official source code for the "Predicting Performance of Heterogeneous AI systems with Discrete-Event Simulations" paper
Basic API, IO, and interfaces of the MinImage system
Задание по курсу параллельного программирования в МФТИ