Skip to content
View KimManjin's full-sized avatar
  • POSTECH
  • South Korea

Highlights

  • Pro

Block or report KimManjin

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results
Python 337 34 Updated Aug 3, 2026

A generalist video MLLM built for fine-grained motion, long-form reasoning, temporal grounding, and online proactive response.

Python 399 8 Updated Jul 27, 2026

TimeLens2: Generalist Video Temporal Grounding with Multimodal LLMs

Python 122 1 Updated Jul 27, 2026

Evaluation code for "Benchmarking Visual State Tracking in Multimodal Video Understanding"

Python 40 1 Updated Aug 12, 2026

Official implementation of EgoHOD at ICLR 2025; 14 EgoVis Challenge Winners in CVPR 2024

Python 37 2 Updated Nov 25, 2025

Official implementation of "HowToCaption: Prompting LLMs to Transform Video Annotations at Scale." ECCV 2024

Python 59 Updated Aug 19, 2025

[ICCV2023] EgoObjects: A Large-Scale Egocentric Dataset for Fine-Grained Object Understanding

Python 86 4 Updated Oct 6, 2023

[Nips 2025] EgoVid-5M: A Large-Scale Video-Action Dataset for Egocentric Video Generation

Python 144 2 Updated Jul 31, 2025

CVPR '26 Highlight

Python 24 Updated May 6, 2026

TIPSv2 (CVPR'26) and TIPS (ICLR'25)

Jupyter Notebook 584 38 Updated Jun 1, 2026

[CVPR 2026 Highlight] A Frame is Worth One Token: Efficient Generative World Modeling with Delta Tokens

Python 232 7 Updated Jul 17, 2026

[ICML 2026] Official implementation of "Deep Forcing: Training-Free Long Video Generation with Deep Sink and Participative Compression"

Python 142 6 Updated Apr 30, 2026

Your own personal AI assistant. Any OS. Any Platform. The lobster way. 🦞

TypeScript 386,808 81,262 Updated Aug 20, 2026

Cosmos Policy

Python 854 99 Updated Jan 23, 2026

Code to pretrain, fine-tune, and evaluate DreamZero and run sim & real-world evals

Python 2,570 225 Updated Apr 19, 2026

Official Python inference and LoRA trainer package for the LTX-2 audio–video generative model.

Python 9,163 1,443 Updated Aug 16, 2026

Frontier Multimodal Foundation Models for Image and Video Understanding

Jupyter Notebook 1,175 88 Updated Aug 14, 2025

Accepted By The 39th Annual Conference on Neural Information Processing Systems Datasets and Benchmarks Track

Python 25 Updated Nov 17, 2025

Official code for MotionBench (CVPR 2025)

Python 76 2 Updated Mar 3, 2025

[NeurIPS 2025 Oral]Infinity⭐️: Unified Spacetime AutoRegressive Modeling for Visual Generation

Python 777 28 Updated Apr 16, 2026

Improving Motion in Image-to-Video Models via Adaptive Low-Pass Guidance (CVPR 2026 Highlight)

Python 60 5 Updated Feb 23, 2026

[RSS 2023] Diffusion Policy Visuomotor Policy Learning via Action Diffusion

Python 4,468 827 Updated Dec 24, 2024

MotionStream: Real-Time Video Generation with Interactive Motion Controls

577 20 Updated Mar 1, 2026

Learning Real-World Action-Video Dynamics with Heterogeneous Masked Autoregression

Python 41 Updated Feb 17, 2025

Official PyTorch implementation of "Video Summarization with Large Language Models" (CVPR 2025).

Python 20 2 Updated Oct 7, 2025

Native Multimodal Models are World Learners

Python 1,549 69 Updated Dec 30, 2025

[ICML2025] The code and data of Paper: Towards World Simulator: Crafting Physical Commonsense-Based Benchmark for Video Generation

Python 165 8 Updated Oct 25, 2024

Official Repo for Self-Forcing++ High Quality Long Video Generation

270 4 Updated Oct 13, 2025

Video Generation, Physical Commonsense, Semantic Adherence, VideoCon-Physics

Python 208 14 Updated Jan 30, 2026

Long Video Gen Infrastructure

Python 2,551 246 Updated Aug 7, 2026
Next