-
NUS, ICT-VIPL
- https://martayang.github.io/
Stars
Unlimited OCR Works: Welcome the Era of One-shot Long-horizon Parsing.
the official implementation of "Compositional Human-Environment Video Synthesis via Spatial-Decoupled Motion Injection and Hybrid Context Integration"
[ICLR 2026] An unified model for 4D human-scene reconstruction
Elevate your AI research writing, no more tedious polishing ✨
Comprehensive open-source library of AI research and engineering skills for any AI model. Package the skills and your claude code/codex/gemini agent will be an AI research agent with full horsepowe…
Codes for the ICCV 2025 paper: "Humans as Checkerboards: Calibrating Camera Motion Scale for World-Coordinate Human Mesh Recovery"
[TPAMI 2025] ViewCrafter: Taming Video Diffusion Models for High-fidelity Novel View Synthesis
Wan: Open and Advanced Large-Scale Video Generative Models
🏋 Modern open-source fitness coaching platform. Create workout plans, track progress, and access a comprehensive exercise database.
[CVPR 2025 Best Paper Award] VGGT: Visual Geometry Grounded Transformer
Metric depth estimation from a single image
Depth Pro: Sharp Monocular Metric Depth in Less Than a Second.
[TMLR 2025🔥] A survey for the autoregressive models in vision.
Code for ICCV 2021 paper "HuMoR: 3D Human Motion Model for Robust Pose Estimation"
Code for the project "MegaSaM: Accurate, Fast and Robust Structure and Motion from Casual Dynamic Videos"
Official Implementation of paper "MonST3R: A Simple Approach for Estimating Geometry in the Presence of Motion"
[ICLR 2026] Official Implementation for Paper "Joint Optimization for 4D Human-Scene Reconstruction in the Wild"
WinSCP is a popular free file manager for Windows supporting SFTP, FTP, FTPS, SCP, S3, WebDAV and local-to-local file transfers. A powerful tool to enhance your productivity with a user-friendly in…
A project page template for academic papers. Demo at https://eliahuhorwitz.github.io/Academic-project-page-template/
The official PyTorch code for RoHM: Robust Human Motion Reconstruction via Diffusion.
High-resolution models for human tasks.
「3D视觉(三维重建、SLAM、AR/VR) + 传统图像处理 + 计算机视觉(偏AI) 」重要知识点和面试问题。
TRAM: Global Trajectory and Motion of 3D Humans from in-the-wild Videos
[CVPR 2024] TokenHMR: Advancing Human Mesh Recovery with a Tokenized Pose Representation
Enhanced ChatGPT Clone: Features Agents, MCP, Skills, DeepSeek, Anthropic, AWS, OpenAI, Responses API, Azure, Groq, o1, GPT-5, Mistral, OpenRouter, Vertex AI, Gemini, Artifacts, AI model switching,…
[IEEE TMM] InstructHumans: Editing Animated 3D Human Textures with Instructions