-
The University of Queensland
- Brisbane
Stars
Skill package for ML/CV/NLP paper writing, curated and adapted from Prof. Peng Sida's open notes for Codex, Claude Code, and Gemini.
Codex Autoresearch Skill — A self-directed iterative system for Codex that continuously cycles through: modify, verify, retain or discard, and repeat indefinitely. Inspired by Karpathy’s autoresear…
An interactive, bilingual literature and evaluation index for world models in multimodal reasoning.
Systematizing Unified Multimodal In-Context Learning through a Capability-Oriented Taxonomy
This repo contains the code for "MEGA-Bench Scaling Multimodal Evaluation to over 500 Real-World Tasks" [ICLR 2025]
[TACL'23] VSR: A probing benchmark for spatial undersranding of vision-language models.
不想啃 5000+ 全文?我已经替你和 LLM 啃完了 — ICLR 2026 全景中文导读
🤗 ml-intern: an open-source ML engineer that reads papers, trains models, and ships ML models
GRU4Rec is the original Theano implementation of the algorithm in "Session-based Recommendations with Recurrent Neural Networks" paper, published at ICLR 2016 and its follow-up "Recurrent Neural Ne…
[WSDM'2024 Oral] "LLMRec: Large Language Models with Graph Augmentation for Recommendation"
Codes for Hierarchical Time-Aware Mixture of Experts for Multi-Modal Sequential Recommendation (WWW2025)
Lightweight coding agent that runs in your terminal
Official code repository for the ACL 2026 paper: "VideoStir: Understanding Long Videos via Spatio-Temporally Structured and Intent-Aware RAG"
Bridging LLM and Recommender System.
Wechat Chat History Exporter 微信聊天记录导出备份程序
AI agents running research on single-GPU nanochat training automatically
MLLMRec-R1: Incentivizing Reasoning Capability in Large Language Models for Multimodal Sequential Recommendation
[ICLR 2025 Oral] Seer: Predictive Inverse Dynamics Models are Scalable Learners for Robotic Manipulation
Helios: Real Real-Time Long Video Generation Model
KDD 2026 | Thought-Augmented Planning for LLM-Powered Interactive Recommender Agent
[ACL2025] "RecLM: Recommendation Instruction Tuning"
Self-evolving Context Database for AI Agents. Unify Agent Memory, Knowledge RAG and Skills.
Source code for NoteLLM and NoteLLM-2
The code repo for our AAAI-25 paper 'Harnessing Multimodal Large Language Models for Multimodal Sequential Recommendation'