Skip to content
View lia112's full-sized avatar

Block or report lia112

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

Post-training with Tinker

Python 4,018 510 Updated Aug 13, 2026

[NeurIPS 2025] An official implementation of Flow-GRPO: Training Flow Matching Models via Online RL

Python 2,466 167 Updated May 7, 2026

[2025] Efficient Vision Language Models: A Survey

52 3 Updated Jul 14, 2025

ARIS ⚔️ (Auto-Research-In-Sleep) — Lightweight Markdown-only skills for autonomous ML research: cross-model review loops, idea discovery, and experiment automation. No framework, no lock-in — works…

Python 14,629 1,290 Updated Aug 11, 2026

Towards Efficient Multimodal Large Language Models: A Survey on Token Compression

216 9 Updated Aug 10, 2026

[Up-to-date] Large Language Model Agent: A Survey on Methodology, Applications and Challenges

2,821 115 Updated Nov 7, 2025

A course in reinforcement learning in the wild

Jupyter Notebook 6,555 1,810 Updated Mar 31, 2026

👀「大模型」2小时从0训练65M参数的视觉多模态VLM!Train a 65M-parameter VLM from scratch in just 2h!

Python 8,461 927 Updated Aug 6, 2026

🧠「大模型」2小时完全从0训练64M的小参数LLM!Train a 64M-parameter LLM from scratch in just 2h!

Python 54,651 7,131 Updated Aug 6, 2026

Unified Efficient Fine-Tuning of 100+ LLMs & VLMs (ACL 2024)

Python 74,056 9,061 Updated Aug 10, 2026

collection of diffusion model papers categorized by their subareas

2,224 102 Updated Mar 16, 2026

T2I-Copilot: A Training-Free Multi-Agent Text-to-Image System for Enhanced Prompt Interpretation and Interactive Generation (ICCV'25)

Jupyter Notebook 57 3 Updated Oct 6, 2025

A compilation of the best multi-agent papers

TeX 1,645 159 Updated Aug 11, 2026

Evolve your language agent with Agentic Context Engineering (ACE)

Python 1,260 162 Updated May 19, 2026

Local UI to run and train LLMs and diffusion models, including Qwen3.8, Kimi K3, MiniMax-H3, Gemma 4, DeepSeek-V4, FLUX and more.

Python 70,733 6,392 Updated Aug 13, 2026

[ICCV'23 Main Track, WECIA'23 Oral] Official repository of paper titled "Self-regulating Prompts: Foundational Model Adaptation without Forgetting".

Python 286 20 Updated Sep 28, 2023

[ICLR 2026] The offical Implementation of "Soft-Prompted Transformer as Scalable Cross-Embodiment Vision-Language-Action Model"

C++ 708 66 Updated Jun 10, 2026

This series will take you on a journey from the fundamentals of NLP and Computer Vision to the cutting edge of Vision-Language Models.

Jupyter Notebook 1,178 104 Updated Jan 23, 2025

EmoBench-M: A benchmark for evaluating Emotional Intelligence in Multimodal Large Language Models (MM 2026)

Python 147 13 Updated Jul 12, 2026

🔥 🔥 🔥 A paper list of some recent Computer Vision(CV) works

967 57 Updated Jun 21, 2026

A curated list of awesome prompt/adapter learning methods for vision-language models like CLIP.

794 42 Updated Jul 17, 2026

A toolbox for skeleton-based action recognition.

Python 1,257 230 Updated Feb 19, 2026

A flexible and extensible framework for gait recognition. You can focus on designing your own models and comparing with state-of-the-arts easily with the help of OpenGait.

Python 1,141 232 Updated Jul 13, 2026

A curated list of Gait Recognition and related area resource

187 13 Updated May 28, 2025

🤗 Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and training.

Python 164,050 34,231 Updated Aug 13, 2026

A PyTorch implementation of the Transformer model in "Attention is All You Need".

Python 9,781 2,099 Updated Apr 16, 2024

Homepage for STAT 157 at UC Berkeley

Jupyter Notebook 4,002 1,508 Updated Feb 16, 2021

deep learning for image processing including classification and object-detection etc.

Python 26,350 8,178 Updated Jan 1, 2026

OpenMMLab Computer Vision Foundation

Python 6,466 1,770 Updated Jan 29, 2026
Next