Stars
Awesome Deep Learning papers for industrial Search, Recommendation and Advertisement. They focus on Embedding, Matching, Pre-Ranking, Ranking, Post Ranking, Relevance, LLM and RL. Please cite our p…
Mobile and Web client for Codex and Claude Code, with realtime voice, encryption and fully featured
[CVPR 2026] EgoMind: Activating Spatial Cognition through Linguistic Reasoning in MLLMs
🔍大模型应用开发实战一:RAG 技术全栈指南,在线阅读地址:https://datawhalechina.github.io/all-in-rag/
Vite & Vue powered static site generator.
🖼️ PNG/JPEG optimization app for macOS, Windows and Linux.
🥢像老乡鸡🐔那样做饭。已添加2026年发布的《老乡鸡菜品溯源报告 2.0中新出现的菜品。主要部分于2024年完工,非老乡鸡官方仓库。文字来自《老乡鸡菜品溯源报告》,并做归纳、编辑与整理。CookLikeHOC.
(2025 AAAI) CoDTS: Enhancing Sparsely Supervised Collaborative Perception with a Dual Teacher-Student Framework
Programmer's guide about how to cook at home.
Learning to Detect Objects from Multi-Agent LiDAR Scans without Manual Labels. (CVPR2025)
Talk to any LLM with hands-free voice interaction, voice interruption, and Live2D taking face running locally across platforms
An AI-powered interactive avatar engine using Live2D, LLM, ASR, TTS, and RVC. Ideal for VTubing, streaming, and virtual assistant applications.
🧠 Train a 64M-parameter LLM from scratch in just 2h!
Best and simplest tool for website change detection, web page monitoring, and website change alerts. Perfect for tracking content changes, price drops, restock alerts, and website defacement monito…
👀 Train a 65M-parameter VLM from scratch in just 2h!
#1 PDF Application on GitHub that lets you edit PDFs on any device anywhere
Valley is a cutting-edge multimodal large model designed to handle a variety of tasks involving text, images, video, and audio data.
Caesium is an image compression software that helps you store, send and share digital pictures, supporting JPG, PNG, WebP and TIFF formats. You can quickly reduce the file size (and resolution, if …
pix2tex: Using a ViT to convert images of equations into LaTeX code.
Qwen3 is the large language model series developed by Qwen team, Alibaba Cloud.
[ICLR2024] HEAL: An Extensible Framework for Open Heterogeneous Collaborative Perception ➡️ All You Need for Multi-Modality Collaborative Perception!
🎨 ML Visuals contains figures and templates which you can reuse and customize to improve your scientific writing.