Skip to content
View jmiemirza's full-sized avatar

Highlights

  • Pro

Block or report jmiemirza

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

TTRV: Test-Time Reinforcement Learning for Vision–Language Models (CVPR 2026)

Python 46 5 Updated Mar 8, 2026

VisualOverload (CVPR 2026) is a VQA benchmark for image understanding in dense, high-resolution scenes.

Python 18 2 Updated May 31, 2026

Repository for the paper: Teaching VLMs to Localize Specific Objects from In-context Examples

Python 40 2 Updated Nov 27, 2024

The repository provides code for running inference with the Meta Segment Anything Model 2 (SAM 2), links for downloading the trained model checkpoints, and example notebooks that show how to use th…

Jupyter Notebook 19,703 2,528 Updated May 30, 2026

ImageBind One Embedding Space to Bind Them All

Python 9,065 845 Updated Nov 21, 2025

Official repository for the MMFM challenge

Python 26 5 Updated Jun 18, 2024

Official inference library for Mistral models

Jupyter Notebook 10,836 1,054 Updated Jun 16, 2026

Video Test-Time Adaptation for Action Recognition (CVPR 2023)

Python 53 5 Updated Oct 13, 2024

Code for the paper: Rotating Features for Object Discovery

Python 53 5 Updated Aug 12, 2024

Cleaner and Formatter for BibTeX files

TeX 1,112 91 Updated Aug 10, 2026

Official PyTorch Implementation of "Scalable Diffusion Models with Transformers"

Python 8,692 809 Updated May 31, 2024

Automated dense category annotation engine that serves as the initial semantic labeling for the Segment Anything dataset (SA-1B).

Python 2,302 142 Updated Jun 7, 2023

Caption-Anything is a versatile tool combining image segmentation, visual captioning, and ChatGPT, generating tailored captions with diverse controls for user preferences. https://huggingface.co/sp…

Python 1,776 104 Updated Aug 29, 2023

SAM with text prompt

Python 2,592 299 Updated Aug 28, 2025

The repository provides code for running inference with the SegmentAnything Model (SAM), links for downloading the trained model checkpoints, and example notebooks that show how to use the model.

Jupyter Notebook 54,683 6,354 Updated Sep 18, 2024

Grounded SAM: Marrying Grounding DINO with Segment Anything & Stable Diffusion & Recognize Anything - Automatically Detect , Segment and Generate Anything

Jupyter Notebook 17,706 1,596 Updated Sep 5, 2024

Official implementation of Cold-Diffusion for different transformations in pytorch.

Python 1,136 83 Updated Oct 13, 2022

Implementation of Denoising Diffusion Probabilistic Model in Pytorch

Python 10,671 1,281 Updated Aug 2, 2026

Unofficial implementation of Palette: Image-to-Image Diffusion Models by Pytorch

Python 1,833 238 Updated Jul 7, 2023

code for deep learning courses

Jupyter Notebook 1,267 334 Updated May 29, 2026

Official PyTorch implementation of ODISE: Open-Vocabulary Panoptic Segmentation with Text-to-Image Diffusion Models [CVPR 2023 Highlight]

Python 945 60 Updated Jul 6, 2024

The simplest, fastest repository for training/finetuning medium-sized GPTs.

Python 62,102 10,702 Updated Nov 12, 2025

EVA Series: Visual Representation Fantasies from BAAI

Python 2,689 187 Updated Aug 1, 2024

Test-time Prompt Tuning (TPT) for zero-shot generalization in vision-language models (NeurIPS 2022))

Python 215 25 Updated Oct 21, 2022

3D point cloud datasets in HDF5 format, containing uniformly sampled 2048 points per shape.

Python 520 54 Updated Sep 20, 2022

pytorch based implementation faster rcnn

Python 435 84 Updated Nov 21, 2020

A minimal PyTorch re-implementation of the OpenAI GPT (Generative Pretrained Transformer) training

Python 24,798 3,314 Updated Aug 15, 2024

Repo for "Benchmarking Robustness of 3D Point Cloud Recognition against Common Corruptions" https://arxiv.org/abs/2201.12296

Python 217 20 Updated Aug 26, 2023

Refine high-quality datasets and visual AI models

TypeScript 11,012 810 Updated Aug 14, 2026

LiDAR snowfall simulation

Python 232 32 Updated Feb 22, 2026
Next