Skip to content
View layumi's full-sized avatar

Block or report layumi

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

From Flatland to Space (SPAR). Accepted to NeurIPS 2025 Datasets & Benchmarks. A large-scale dataset & benchmark for 3D spatial perception and reasoning in VLMs.

Python 94 2 Updated Jan 5, 2026

[CVPR 2026 Oral] VGGT Omega

Python 3,972 274 Updated Jul 15, 2026

🚀Official Repository of Intelligent Remote Sensing Agents: A Survey

594 14 Updated Jul 29, 2026

Exclusively Dark (ExDARK) dataset which to the best of our knowledge, is the largest collection of low-light images taken in very low-light environments to twilight (i.e 10 different conditions) to…

MATLAB 640 114 Updated Feb 13, 2026

[ICLR 2026] π^3: Permutation-Equivariant Visual Geometry Learning

Python 2,107 166 Updated Jul 3, 2026

[ICLR 2026] Sat3DGen: Comprehensive Street-Level 3D Scene Generation from Single Satellite Image

Python 122 10 Updated Aug 4, 2026
3 Updated Apr 1, 2026

🚁 Can Vision-Language Models Think from the Sky? UAVReason for Aerial Reasoning and Generation

Python 23 Updated Jul 11, 2026

Road Maps as Free Geometric Priors: Weather-Invariant Drone Geo-Localization with GeoFuse

17 Updated Jul 6, 2026

[ACL 2026] Code for "Cultivating Forensic Reasoning for Generalizable Multimodal Manipulation Detection"

Python 13 Updated Jun 13, 2026

[ACL 2026 Main] Official implementation of "Generating Attribution Reports for Manipulated Facial Images"

Python 3 Updated Apr 20, 2026

The official code of "Pretrain-then-Adapt: Uncertainty-Aware Test-Time Adaptation for Text-based Person Search" [SIGIR 2026]

Python 7 Updated Jul 22, 2026

(ACM TOMM) This is the official code repository for "VM-UNet: Vision Mamba UNet for Medical Image Segmentation".

Python 860 57 Updated Sep 3, 2025

[PR 2026] Harnessing Weak Pair Uncertainty for Text-based Person Search

2 Updated Apr 12, 2026

[IEEE Transactions on Image Processing'26] Pytorch implementation of FANet: Fovea Attention Network for Robust Aerial Geo-localization Across Diverse Weather Conditions

Python 5 Updated Jun 22, 2026

[ACL 2026] Hard to Read, Easy to Jailbreak: How Visual Degradation Bypasses MLLM Safety Alignment

Python 13 1 Updated May 11, 2026

[ACL 2024 Findings] MedAgents: Large Language Models as Collaborators for Zero-shot Medical Reasoning https://arxiv.org/abs/2311.10537

Python 364 50 Updated May 27, 2024

The official implementation of Error Detection in Egocentric Procedural Task Videos

Python 36 6 Updated Sep 20, 2025

Graph-Backed Generative Brick Assembly

Python 36 3 Updated Jul 7, 2026
29 Updated Jun 29, 2026

SUES-200: A Multi-height Multi-scene Cross-view Image Benchmark Across Drone and Satellite

Python 90 3 Updated Nov 6, 2024

[NeurIPS 2024] This repo contains evaluation code for the paper "Are We on the Right Way for Evaluating Large Vision-Language Models"

Python 215 5 Updated Sep 26, 2024

A Simple and Universal Swarm Intelligence Engine, Predicting Anything. 简洁通用的群体智能引擎,预测万物

Python 70,942 11,050 Updated Aug 3, 2026

TurtleBench: Evaluating Top Language Models via Real-World Yes/No Puzzles.

Jupyter Notebook 161 15 Updated Oct 16, 2024

PyTorch building blocks for the OLMo ecosystem

Python 1,467 304 Updated Aug 14, 2026

Modeling, training, eval, and inference code for OLMo

Python 6,625 790 Updated Nov 24, 2025

Grounded SAM: Marrying Grounding DINO with Segment Anything & Stable Diffusion & Recognize Anything - Automatically Detect , Segment and Generate Anything

Jupyter Notebook 17,708 1,596 Updated Sep 5, 2024

全维度的前沿大语言模型自动化评测套件。涵盖逻辑推理、智能体编程、网页特效代码生成以及百万Token级长文本解析(GPT-5.4 / Claude 4.7 / DeepSeek-V4 等)

HTML 45 8 Updated Apr 29, 2026
Next