Lists (2)
Sort Name ascending (A-Z)
Stars
The largest collection of PyTorch image encoders / backbones. Including train, eval, inference, export scripts, and pretrained weights -- ResNet, ResNeXT, EfficientNet, NFNet, Vision Transformer (V…
The repository for 3D Vision Course Project 'Video Depth Estimation in 2025'
EPNet++: Cascade Bi-directional Fusion for Multi-Modal 3D Object Detection (TPAMI-2022)
EPNet: Enhancing Point Features with Image Semantics for 3D Object Detection(ECCV 2020)
A single stage range image based 3D object detector
Official Pytorch implementation of Birdnet+
Paper and Codes for “RangeDet: In Defense of Range View for LiDAR-based 3D Object Detection” (ICCV2021)
Transferable Semi-supervised 3D Object Detection from RGB-D Data, ICCV 2019
TensorRT deploy and PTQ/QAT tools development for FastBEV, total time only need 6.9ms!!!
Fast-BEV: A Fast and Strong Bird’s-Eye View Perception Baseline
OpenMMLab's next-generation platform for general 3D object detection.
[NeurIPS 2022 & TPAMI 2025] DeepInteraction: 3D Object Detection via Modality Interaction
Lift, Splat, Shoot: Encoding Images from Arbitrary Camera Rigs by Implicitly Unprojecting to 3D (ECCV 2020)
Offical PyTorch implementation of "BEVFusion: A Simple and Robust LiDAR-Camera Fusion Framework"
This repository is an open-source PointPainting package which is easy to understand, deploy and run!
[CVPR2021] PointAugmenting: Cross-Modal Augmentation for 3D Object Detection
[ICCV 2023] Cross Modal Transformer: Towards Fast and Robust 3D Object Detection
[ICRA 2024] SuperFusion: Multilevel LiDAR-Camera Fusion for Long-Range HD Map Generation
[ECCV2022, IJCAI2022] AutoAlignV2: Deformable Feature Aggregation for Dynamic Multi-Modal 3D Object Detection
MetaBEV: Solving Sensor Failures for BEV Detection and Map Segmentation
[IV'24] UniBEV: the official implementation of UniBEV
[ECCV2024] This is the official implementation of GraphBEV, a BEV multi-modal framework for autonomous driving perception, e.g., 3D object detection and semantic map segmentation.
[ICRA'23] BEVFusion: Multi-Task Multi-Sensor Fusion with Unified Bird's-Eye View Representation
Deep Learning Book Chinese Translation
Interactive deep learning book with multi-framework code, math, and discussions. Adopted at 500 universities from 70 countries including Stanford, MIT, Harvard, and Cambridge.
CVT-Occ: Cost Volume Temporal Fusion for 3D Occupancy Prediction