Stars
Visualize, query, and stream to train on multimodal robotics data.
Official Implementation of paper "MonST3R: A Simple Approach for Estimating Geometry in the Presence of Motion"
cubao / Open3D
Forked from isl-org/Open3DSee `releases` for binary wheels for ubuntu 16.04
Official implementation of Hierarchical Context Merging: Better Long Context Understanding for Pre-trained LLMs (ICLR 2024).
A python tool for fitting primitives 3D shapes in point clouds using RANSAC algorithm
This repository contains information for the paper "A Survey on RGB-D Datasets" and is a collaborative initiative to update the datasets list faster.
Jaehyung Kim et al's ICML23 paper "Prefer to Classify: Improving Text Classifiers via Auxiliary Preference Learning"
Jaehyung Kim et al's ACL 2023 paper on "infoVerse: A Universal Framework for Dataset Characterization with Multidimensional Meta-information"
The repository provides code for running inference with the SegmentAnything Model (SAM), links for downloading the trained model checkpoints, and example notebooks that show how to use the model.
Python Script to download hundreds of images from 'Google Images'. It is a ready-to-run code!
IFSeg: Image-free Semantic Segmentation via Vision-Language Model (CVPR 2023)
[CVPR'22 Oral] Temporal Alignment Networks for Long-term Video. Tengda Han, Weidi Xie, Andrew Zisserman.
[NeurIPS 2022 Spotlight] VideoMAE: Masked Autoencoders are Data-Efficient Learners for Self-Supervised Video Pre-Training
[ICLR2022] official implementation of UniFormer
[ICLR'22] Generating Videos with Dynamics-aware Implicit Generative Adversarial Networks
Meta-Learning Sparse Implicit Neural Representations (NeurIPS 2021)
Implementation / replication of DALL-E, OpenAI's Text to Image Transformer, in Pytorch
Code for "Compositional Video Synthesis with Action Graphs", Bar & Herzig et al., ICML 2021
PyTorch code for Vision Transformers training with the Self-Supervised learning method DINO
OpenMMLab's Next Generation Video Understanding Toolbox and Benchmark
COOT: Cooperative Hierarchical Transformer for Video-Text Representation Learning
Lightweight, dependency-free Python library and CLI for downloading YouTube videos, playlists, and captions.
wsshin / jemdoc_mathjax
Forked from jem/jemdocjemdoc with MathJax support and more
Tensors and Dynamic neural networks in Python with strong GPU acceleration
Retrieve author and publication information from Google Scholar in a friendly, Pythonic way without having to worry about CAPTCHAs!