Skip to content
View wd1511's full-sized avatar
🎯
Focusing
🎯
Focusing
  • Harbin Institute of Technology
  • Shanghai, China
  • 16:03 (UTC +08:00)

Block or report wd1511

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

[CVPR 2026] "GenieDrive: Towards Physics-Aware Driving World Model with 4D Occupancy Guided Video Generation"

Python 113 4 Updated May 18, 2026

[ACMMM 2025] Officially implement of the paper "DriVerse: Navigation World Model for Driving Simulation via Multimodal Trajectory Prompting and Motion Alignment"

Python 222 3 Updated May 7, 2025

Official implementation of paper - PnLCalib: Sports Field Registration via Points and Lines Optimization

Python 104 16 Updated Mar 17, 2026

Official Implementation of "DriveWAM: Video Generative Priors Enable Scalable World-Action Modeling for Autonomous Driving"

Python 54 1 Updated Jul 19, 2026

NVIDIA Cosmos-Dreams (fka NVIDIA OmniDreams) is a world model that generates photorealistic video for autonomous-driving simulation in real time.

Python 304 23 Updated Jul 24, 2026

[ICCV-2025 Spotlight] Official implementation of SEGA: A stepwise evolution paradigm for content-aware layout generation with design prior

Python 51 5 Updated Mar 25, 2026

MiniMax LLM + Pi05 VLA Robot Agent Demo

Python 58 8 Updated Dec 31, 2025

Official repo for paper "Echo-Infinity: Learnable Evolving Memory for Real-Time Infinite Video Generation"

Python 106 3 Updated Jun 4, 2026

[CVPR 2025 Best Paper Award] VGGT: Visual Geometry Grounded Transformer

Python 14,229 1,533 Updated May 19, 2026

🌐 Event Camera Vision in the Era of Large Models: A Survey

TeX 10 2 Updated Jul 28, 2026

当下热门的模糊人脸修复模型的部署,分别是:Codeformer,GFPGAN,GPEN,Restoreformer

Python 42 10 Updated Sep 14, 2023

Learning Compressed Representation of 3DLUT for Image-enhancement. Higher performance with much smaller models! ☺️

Python 141 7 Updated Jan 10, 2024

Official repository of "BasicVSR++: Improving Video Super-Resolution with Enhanced Propagation and Alignment"

Python 818 82 Updated Dec 27, 2023
Jupyter Notebook 2,612 454 Updated Mar 10, 2026

A Collection of Papers and Codes for CVPR2026/CVPR2025/CVPR2024/CVPR2021/CVPR2020 Low Level Vision

1,733 168 Updated Jun 26, 2026

Paper list for video enhancement, including video super-resolution, interpolation, denoising, deblurring and inpainting.

156 6 Updated Aug 28, 2024

KBNet: Kernel Basis Network for Image Restoration

Python 218 32 Updated Oct 13, 2025

The official PyTorch implementation for CascadedGaze: Efficiency in Global Context Extraction for Image Restoration, TMLR'24.

Python 80 5 Updated Feb 13, 2025

Image Restoration Toolbox (PyTorch). Training and testing codes for DPIR, USRNet, DnCNN, FFDNet, SRMD, DPSR, BSRGAN, SwinIR

Python 3,519 706 Updated Oct 2, 2024

A PyTorch implementation of "Real-time Scene Text Detection with Differentiable Binarization".

Python 2,261 486 Updated Mar 11, 2024

A PyTorch implementation of "TextFuseNet: Scene Text Detection with Richer Fused Features".

Python 483 122 Updated Jul 2, 2021

Implementation of "Rapid Salient Object Detection with Difference Convolutional Neural Networks"

Python 26 4 Updated Jul 23, 2025

The code for our newly accepted paper in Pattern Recognition 2020: "U^2-Net: Going Deeper with Nested U-Structure for Salient Object Detection."

Python 9,848 1,619 Updated Jun 26, 2024

Deep Contextual Video Compression

Python 809 129 Updated Aug 7, 2026

Official implementation of Character Region Awareness for Text Detection (CRAFT)

Python 3,395 927 Updated Jul 16, 2024

Turn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/PDFs and LLMs. Supports 100+ languages.

Python 87,623 11,176 Updated Jul 22, 2026

Easy-to-use image segmentation library with awesome pre-trained model zoo, supporting wide-range of practical tasks in Semantic Segmentation, Interactive Segmentation, Panoptic Segmentation, Image …

Python 9,373 1,710 Updated Feb 5, 2026

Object Detection toolkit based on PaddlePaddle. It supports object detection, instance segmentation, multiple object tracking and real-time multi-person keypoint detection.

Python 14,374 3,025 Updated May 28, 2026
Next