-
Tsinghua University
-
19:17
(UTC +08:00) - https://orcid.org/0000-0003-2333-7493
Stars
A curated list which collects the latest advance to accelerate VGGT
A collection of token reduction (token pruning, merging, clustering, etc.) techniques for ML/AI
[ICLR 2026] 🐻 Uniform Discrete Diffusion with Metric Path for Video Generation
VisionSelector: End-to-End Learnable Visual Token Compression for Efficient Multimodal LLMs
Estimating the inference time of pytorch models on the Android platform
👉 CARLA resources such as tutorial, blog, code and etc https://github.com/carla-simulator/carla
Awesome LLM compression research papers and tools.
[ICCV 2025] SAM4D: Segment Anything in Camera and LiDAR Streams
LSTS: Periodicity Learning via Long Short-term Temporal Shift for Remote Physiological Measurement
DSPy: The framework for programming—not prompting—language models
Customizable Multimodal Trajectory Prediction via Nodes of Interest Selection for Autonomous Vehicles
AutoLibra: Metric Induction for Agents from Open-Ended Human Feedback
Lets make video diffusion practical!
OpenUI let's you describe UI using your imagination, then see it rendered live.
分享一些好用的 Dify DSL 工作流程,自用、学习两相宜。 Sharing some Dify workflows.
AI video translation & dubbing tool for humans and AI Agents, powered by LLMs. Full pipeline: download, transcribe, translate, TTS dub, reformat, cover generation. 100+ languages, optimized for You…
Build Real-Time Knowledge Graphs for AI Agents
About Awesome things towards foundation agents. Papers / Repos / Blogs / ...
Awesome-llm-role-playing-with-persona: a curated list of resources for large language models for role-playing with assigned personas
Code and Data for EMNLP 2024 Paper "Neeko: Leveraging Dynamic LoRA for Efficient Multi-Character Role-Playing Agent"
[ICLR 2025] Pad: Personalized alignment of llms at decoding-time
Code for the paper "CoS: Enhancing Personalization and Mitigating Bias with Context Steering"
A Survey on Large Language Model-Based Game Agents (ACM CSUR)
[AAAI 2026] OpenDriveVLA: Towards End-to-end Autonomous Driving with Large Vision Language Action Model
[TMLR 2025] Stop Overthinking: A Survey on Efficient Reasoning for Large Language Models