-
Samsung Electronics
- Seoul, Korea
Stars
Lightweight coding agent that runs in your terminal
A collection of original, innovative ideas and algorithms towards Advanced Literate Machinery. This project is maintained by the OCR Team in the Language Technology Lab, Tongyi Lab, Alibaba Group.
Official code of "EVF-SAM: Early Vision-Language Fusion for Text-Prompted Segment Anything Model"
❄️ ChatGPT Desktop Application (Mac, Windows and Linux)
Grounded SAM: Marrying Grounding DINO with Segment Anything & Stable Diffusion & Recognize Anything - Automatically Detect , Segment and Generate Anything
Learning from synthetic data - code and models
Official repository for "AM-RADIO: Reduce All Domains Into One"
Official PyTorch implementation of "Weakly Supervised Semantic Segmentation for Driving Scenes", AAAI2024
[WACV 2023] Image Completion with Heterogeneously Filtered Spectral Hints
Harnessing Large Language Models for Planning: A Lab on Strategies for Success and Mitigation of Pitfalls @ AAAI-24
[ICML 2024] Vision Mamba: Efficient Visual Representation Learning with Bidirectional State Space Model
Official code for "Diversity-Measurable Anomaly Detection", CVPR 2023 by Wenrui Liu, Hong Chang, Bingpeng Ma, Shiguang Shan, Xilin Chen.
ICCV 2023-2025 Papers: Discover cutting-edge research from ICCV 2023-25, the leading computer vision conference. Stay updated on the latest in computer vision and deep learning, with code included.…
The official implementation of VLPL: Vision Language Pseudo Label for Multi-label Learning with Single Positive Labels
An Attribute-based Method for Video Anomaly Detection (TMLR 2025)
A curated list of papers and code in exploring single positive multi-label learning (SPML), a interesting and challenging variant of multi-label learning.
COYO-700M: Large-scale Image-Text Pair Dataset
Tutorials for FLAVA model https://arxiv.org/abs/2112.04482
[CVPR 2022] Official code for "Unified Contrastive Learning in Image-Text-Label Space"
Large-scale Self-supervised Pre-training Across Tasks, Languages, and Modalities
[CVPR 2023] Bridging the Gap between Model Explanations in Partially Annotated Multi-label Classification
Implementation for "DualCoOp: Fast Adaptation to Multi-Label Recognition with Limited Annotations" (NeurIPS 2022))
[NeurIPS 2023] Official implementation of the paper "Segment Everything Everywhere All at Once"
Automated dense category annotation engine that serves as the initial semantic labeling for the Segment Anything dataset (SA-1B).