-
Harbin Institute of Technology
- Shanghai, China
-
16:03
(UTC +08:00) - wd1511.github.io
Lists (2)
Sort Name ascending (A-Z)
Stars
[CVPR 2026] "GenieDrive: Towards Physics-Aware Driving World Model with 4D Occupancy Guided Video Generation"
[ACMMM 2025] Officially implement of the paper "DriVerse: Navigation World Model for Driving Simulation via Multimodal Trajectory Prompting and Motion Alignment"
Official implementation of paper - PnLCalib: Sports Field Registration via Points and Lines Optimization
Official Implementation of "DriveWAM: Video Generative Priors Enable Scalable World-Action Modeling for Autonomous Driving"
NVIDIA Cosmos-Dreams (fka NVIDIA OmniDreams) is a world model that generates photorealistic video for autonomous-driving simulation in real time.
[ICCV-2025 Spotlight] Official implementation of SEGA: A stepwise evolution paradigm for content-aware layout generation with design prior
MiniMax LLM + Pi05 VLA Robot Agent Demo
Official repo for paper "Echo-Infinity: Learnable Evolving Memory for Real-Time Infinite Video Generation"
[CVPR 2025 Best Paper Award] VGGT: Visual Geometry Grounded Transformer
🌐 Event Camera Vision in the Era of Large Models: A Survey
当下热门的模糊人脸修复模型的部署,分别是:Codeformer,GFPGAN,GPEN,Restoreformer
Learning Compressed Representation of 3DLUT for Image-enhancement. Higher performance with much smaller models!
Official repository of "BasicVSR++: Improving Video Super-Resolution with Enhanced Propagation and Alignment"
A Collection of Papers and Codes for CVPR2026/CVPR2025/CVPR2024/CVPR2021/CVPR2020 Low Level Vision
Paper list for video enhancement, including video super-resolution, interpolation, denoising, deblurring and inpainting.
KBNet: Kernel Basis Network for Image Restoration
The official PyTorch implementation for CascadedGaze: Efficiency in Global Context Extraction for Image Restoration, TMLR'24.
Image Restoration Toolbox (PyTorch). Training and testing codes for DPIR, USRNet, DnCNN, FFDNet, SRMD, DPSR, BSRGAN, SwinIR
A PyTorch implementation of "Real-time Scene Text Detection with Differentiable Binarization".
A PyTorch implementation of "TextFuseNet: Scene Text Detection with Richer Fused Features".
Implementation of "Rapid Salient Object Detection with Difference Convolutional Neural Networks"
The code for our newly accepted paper in Pattern Recognition 2020: "U^2-Net: Going Deeper with Nested U-Structure for Salient Object Detection."
Official implementation of Character Region Awareness for Text Detection (CRAFT)
Turn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/PDFs and LLMs. Supports 100+ languages.
Easy-to-use image segmentation library with awesome pre-trained model zoo, supporting wide-range of practical tasks in Semantic Segmentation, Interactive Segmentation, Panoptic Segmentation, Image …
Object Detection toolkit based on PaddlePaddle. It supports object detection, instance segmentation, multiple object tracking and real-time multi-person keypoint detection.