Skip to content
View leotac's full-sized avatar

Highlights

  • Pro

Block or report leotac

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

An efficient video loader for deep learning with smart shuffling that's super easy to digest

C++ 64 9 Updated May 30, 2026

VTGNet: A Vision-based Trajectory Generation Network for Autonomous Vehicles in Urban Environments

Python 43 12 Updated Nov 27, 2020

See the Future: A Semantic Segmentation Network Predicting Ego-vehicle Trajectory with a Single Monocular Camera

8 1 Updated May 27, 2021

[CVPR 2025 Best Paper Award] VGGT: Visual Geometry Grounded Transformer

Python 14,232 1,533 Updated May 19, 2026

RT-GENE: Real-Time Eye Gaze and Blink Estimation in Natural Environments

Python 444 73 Updated Jul 14, 2026

The official PyTorch implementation of L2CS-Net for gaze estimation and tracking

Python 520 113 Updated Feb 2, 2024

This repository contains demos I made with the Transformers library by HuggingFace.

Jupyter Notebook 11,744 1,738 Updated Apr 20, 2026

Dual Swin Transformer for video-time-series fusion

Python 19 4 Updated Aug 28, 2024

Access large language models from the command-line

Python 12,363 950 Updated Aug 12, 2026

Recipes for shrinking, optimizing, customizing cutting edge vision models. 💜

Jupyter Notebook 1,973 152 Updated Aug 12, 2026

Refine high-quality datasets and visual AI models

TypeScript 11,015 810 Updated Aug 14, 2026

Official repo for our paper: "What Matters in Autonomous Driving Anomaly Detection: A Weakly Supervised Horizon"

2 Updated Oct 21, 2024
Python 638 70 Updated Jun 7, 2025

Annotation for reproducing the result of the paper "Cross-model temporal cooperation via saliency maps for efficient recognition and classification of relevant traffic lights" .

Python 6 Updated Sep 28, 2024

Software Development Kit for the Zenseact Open Dataset (ZOD)

Python 147 19 Updated Aug 10, 2026

A playbook for systematically maximizing the performance of deep learning models.

30,280 2,420 Updated Jun 18, 2024

The simplest, fastest repository for training/finetuning medium-sized GPTs.

Python 62,108 10,705 Updated Nov 12, 2025

Code for robust monocular depth estimation described in "Ranftl et. al., Towards Robust Monocular Depth Estimation: Mixing Datasets for Zero-shot Cross-dataset Transfer, TPAMI 2022"

Python 5,424 725 Updated Aug 23, 2024

A latent text-to-image diffusion model

Jupyter Notebook 73,300 10,577 Updated Jun 18, 2024

Render JSON into collapsible HTML

JavaScript 435 92 Updated Mar 7, 2023

[CVPR2023] The official repo for OC-SORT: Observation-Centric SORT on video Multi-Object Tracking. OC-SORT is simple, online and robust to occlusion/non-linear motion.

Python 1,128 153 Updated Apr 21, 2026

[ECCV 2022] ByteTrack: Multi-Object Tracking by Associating Every Detection Box

Python 6,628 1,140 Updated Jun 19, 2024

Awesome list for research on CLIP (Contrastive Language-Image Pre-Training).

1,227 58 Updated Jun 28, 2024

[ICCV 2021] Deep Reinforced Accident Anticipation with Visual Explanation

Python 31 8 Updated Sep 4, 2023

GLIDE: a diffusion-based text-conditional image synthesis model

Python 3,687 499 Updated Mar 8, 2024

[ACM MM 2020] CCD dataset for traffic accident anticipation.

157 13 Updated Sep 2, 2023

Optimization Models used in my e-book with the same title

HTML 14 2 Updated Dec 17, 2025

This is the repo for our Detection of Traffic Anomaly (DoTA) dataset.

Python 272 38 Updated Dec 28, 2023

[ECCV 2020] Learning stereo from single images using monocular depth estimation networks

Python 408 56 Updated Jul 2, 2021
Next