Stars
Code for ICCV2019 "Symmetric Cross Entropy for Robust Learning with Noisy Labels"
[CVPR 2026] The official PyTorch implementation of the "Vision Transformer Needs More Than Registers".
LLM Architecture Gallery source data
A visual AI assistant powered by Qwen3.5 for browser automation
[ICLR 2026] An official implementation of "SIM-CoT: Supervised Implicit Chain-of-Thought"
Youtu-Embedding is an industry-leading, general-purpose text representation model developed by Tencent Youtu Lab.
REFRAG-style RAG (compress → sense/select → expand) — Single-file reference implementation
Spherical Merge Pytorch/HF format Language Models with minimal feature loss.
🧠「大模型」2小时完全从0训练64M的小参数LLM!Train a 64M-parameter LLM from scratch in just 2h!
slime is an LLM post-training framework for RL Scaling.
Awesome Deep Learning papers for industrial Search, Recommendation and Advertisement. They focus on Embedding, Matching, Pre-Ranking, Ranking, Post Ranking, Relevance, LLM and RL. Please cite our p…
PyTorch implementation of the NIPS-17 paper "Poincaré Embeddings for Learning Hierarchical Representations"
A curated list of awesome prompt/adapter learning methods for vision-language models like CLIP.
Official PyTorch Code for Anchor Token Guided Prompt Learning Methods: [ICCV 2025] ATPrompt and [Arxiv 2511.21188] AnchorOPT
Official Code for "Painting with Words: Elevating Detailed Image Captioning with Benchmark and Alignment Learning" (ICLR 2025)
A blazing fast inference solution for text embeddings models
[ICLR 2025 Oral] "Your Mixture-of-Experts LLM Is Secretly an Embedding Model For Free"
The project pages for LexVLA (Unified Lexical Representation for Interpretable Visual-Language Alignment).
Advanced AI Explainability for computer vision. Support for CNNs, Vision Transformers, Classification, Object detection, Segmentation, Image similarity and more.
The repository provides code for running inference with the Meta Segment Anything Model 2 (SAM 2), links for downloading the trained model checkpoints, and example notebooks that show how to use th…
X-VLM: Multi-Grained Vision Language Pre-Training (ICML 2022)
🍀 Pytorch implementation of various Attention Mechanisms, MLP, Re-parameter, Convolution, which is helpful to further understand papers.⭐⭐⭐