Skip to content
View Hongje's full-sized avatar

Block or report Hongje

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

RynnValue: Scaling Robotic Value Foundation Models with Temporal Distance

Python 26 2 Updated Aug 11, 2026

Official Implementation of SCALE: Self-uncertainty Conditioned Adaptive Looking and Execution for Vision-Language-Action Models (ICML'26 Spotlight)

Python 17 2 Updated Jun 29, 2026

TurboVLA: Real-Time Vision-Language-Action Model at 32 Hz on an RTX 4090 with <1 GB VRAM

Python 436 48 Updated Aug 5, 2026

ITS3D: Inference-Time Scaling for Text-Guided 3D Diffusion Models

Jupyter Notebook 11 1 Updated Jul 30, 2026

A general framework for inference-time scaling and steering of diffusion models with arbitrary rewards.

Jupyter Notebook 232 22 Updated Jun 26, 2025
Jupyter Notebook 511 40 Updated Jul 11, 2023

Official repo for Lookahead Sample Reward Guidance for Test-Time Scaling of Diffusion Models (ICML 2026 Spotlight)

Python 5 Updated May 2, 2026

Generalizable VLA Finetuning via Representation Anchoring and Language-Action Alignment

Python 23 1 Updated Jul 21, 2026
Python 945 90 Updated Jun 29, 2026

[ECCV 2026] WildWorld: A Large-Scale Dataset for Dynamic World Modeling with Actions and Explicit State toward Generative ARPG

Python 419 5 Updated Jul 20, 2026

LLM-as-a-Verifier is a general-purpose framework that provides fine-grained feedback for any agent without requiring additional training. It achieves SOTA performance across coding, robotics, and m…

Python 635 57 Updated Aug 7, 2026

PixWorld: Unifying 3D Scene Generation and Reconstruction in Pixel Space

242 10 Updated Jul 24, 2026

Code for <Environmental Change Detection for Real-World Change Analysis> in ECCV 2026

2 Updated Jun 25, 2026

Multi-Resolution Flow Matching: Training-Free Diffusion Acceleration via Staged Sampling

Python 283 18 Updated Jul 24, 2026

Code for 'Not All Prediction Targets Keep Training-Free Diffusion Guidance on the Manifold' (ECCV 2026)

Python 3 Updated Aug 4, 2026

[ECCV 2026] SAM2Matting: Generalized Image and Video Matting

Python 117 8 Updated Jun 30, 2026

From a single casual image to a visually consistent and physically stable interactive 3D scene.

Python 257 19 Updated Jun 23, 2026

TriSplat: Simulation-Ready Feed-Forward 3D Scene Reconstruction

Python 348 23 Updated Jun 12, 2026

[CVPR 2026] VLM-3R: Vision-Language Models Augmented with Instruction-Aligned 3D Reconstruction

Python 437 29 Updated Jul 15, 2026

[CVPR 2026] G2VLM: Geometry Grounded Vision Language Model with Unified 3D Reconstruction and Spatial Reasoning

Python 349 14 Updated Apr 18, 2026

PhysX-Anything: Simulation-Ready Physical 3D Assets from Single Image (CVPR 2026)

Jupyter Notebook 920 60 Updated Apr 28, 2026

GenClaw: Code-Driven Agentic Image Generation

Python 304 7 Updated Jul 19, 2026

NVIDIA Cosmos is an open platform of world models, datasets, and tools that enables developers to build Physical AI for robots, autonomous vehicles, smart infrastructure, and more.

Jupyter Notebook 11,484 823 Updated Aug 13, 2026
Python 1 Updated Feb 2, 2026

The official repository of Qwen-VLA

738 26 Updated May 29, 2026

Official implementation of paper "VLM³: Vision Language Models Are Native 3D Learners".

Jupyter Notebook 416 12 Updated Jul 26, 2026

A curated list of awesome 3D object and scene generation papers.

26 1 Updated May 31, 2026

[CVPR2026] Detect Anything via Next Point Prediction

Jupyter Notebook 1,547 110 Updated Feb 22, 2026
Next