Skip to content
View AadSah's full-sized avatar
🇮🇳
🇮🇳

Organizations

@CVIR

Block or report AadSah

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

Domain-agnostic multi-agent software design and evolution harness

Python 54 21 Updated Jul 25, 2026

Cosmos-Reason2 models understand the physical common sense and generate appropriate embodied decisions in natural language through long chain-of-thought reasoning processes.

Python 432 90 Updated Jun 7, 2026

This repository contains the code for the paper - "Conversational Image Segmentation: Grounding Abstract Concepts with Scalable Supervision" (CVPR 2026)

Python 45 2 Updated Jun 3, 2026

EfficientSAM3 compresses SAM3 into lightweight, edge-friendly models via progressive knowledge distillation for fast promptable concept segmentation and tracking.

Jupyter Notebook 641 49 Updated Jul 1, 2026

Stereo4D dataset and processing code

Jupyter Notebook 310 11 Updated Nov 4, 2025

Unofficial DynaDUSt3R reimplementation trained on Stereo4D (research only).

Python 61 3 Updated Oct 18, 2025

NextFlow🚀: Unified Sequential Modeling Activates Multimodal Understanding and Generation

331 15 Updated Jan 9, 2026

[ICLR 2026] Training Visual Reasoners with Multimodal Verifiers

Jupyter Notebook 14 Updated Jun 12, 2026

State-of-the-art Image & Video CLIP, Multimodal Large Language Models, and More!

Jupyter Notebook 2,329 157 Updated Apr 13, 2026

[arXiv 2023] Set-of-Mark Prompting for GPT-4V and LMMs

Python 1,549 111 Updated Aug 19, 2024

[NeurIPS 2023] Official implementation of the paper "Segment Everything Everywhere All at Once"

Python 4,795 452 Updated Aug 19, 2024

Official repository of paper "Subobject-level Image Tokenization" (ICML-25)

Python 94 10 Updated Jul 4, 2025

Official inference framework for 1-bit LLMs

C++ 39,786 3,654 Updated Jul 21, 2026

Repair malformed JSON from LLMs, APIs, logs, and user input in Python.

Python 5,045 206 Updated Jul 21, 2026

Everything about the SmolLM and SmolVLM family of models

Python 3,853 300 Updated May 26, 2026

A suite of image and video neural tokenizers

Jupyter Notebook 1,732 90 Updated Feb 11, 2025

Hello AI World guide to deploying deep-learning inference networks and deep vision primitives with TensorRT and NVIDIA Jetson.

C++ 8,939 3,102 Updated Oct 16, 2025

Densely Captioned Images (DCI) dataset repository.

Python 197 5 Updated Jul 1, 2024
Jupyter Notebook 205 12 Updated Apr 21, 2026

Grounded SAM 2: Ground and Track Anything in Videos with Grounding DINO, Florence-2 and SAM 2

Jupyter Notebook 3,655 425 Updated Nov 11, 2025

Grounded SAM: Marrying Grounding DINO with Segment Anything & Stable Diffusion & Recognize Anything - Automatically Detect , Segment and Generate Anything

Jupyter Notebook 17,686 1,594 Updated Sep 5, 2024

Caption-Anything is a versatile tool combining image segmentation, visual captioning, and ChatGPT, generating tailored captions with diverse controls for user preferences. https://huggingface.co/sp…

Python 1,775 104 Updated Aug 29, 2023

Waymo Open Dataset

Python 3,372 699 Updated Jan 8, 2026

Data release for the ImageInWords (IIW) paper.

JavaScript 224 8 Updated Nov 17, 2024

Referring Expression Datasets API

Jupyter Notebook 573 85 Updated Aug 27, 2024

[IJCV 2026] Multimodal Referring Segmentation

254 4 Updated Jun 30, 2026

code for affordance-r1

Python 73 4 Updated May 11, 2026

Project Page for "LISA: Reasoning Segmentation via Large Language Model"

Python 2,667 208 Updated Feb 16, 2025

Reference PyTorch implementation and models for DINOv3

Jupyter Notebook 11,006 910 Updated Jul 15, 2026

Official Repository for MolmoAct

Python 376 42 Updated May 11, 2026
Next