-
UESTC | UniTn
- Sichuan ⇌ Italy
-
08:40
(UTC +08:00) - https://zchoi.github.io/
Highlights
- Pro
-
-
-
This is a curated list of "Embodied AI or robot with Large Language Models" research. Watch this repository for the latest updates! 🔥
-
zchoi.github.io Public
Forked from RayeRen/acad-homepage.github.ioAcadHomepage: A Modern and Responsive Academic Personal Homepage
-
MMPoT-Survey Public
Project website for A Survey on Post-training of Multimodal Large Language Models.
-
A curated list of "A Survey on Post-training of Multimodal Large Language Models" research. Watch this repository for latest updates! 🔥
-
-
OmniCharacter-plus Public
[TPAMI26] Official codebase for "OmniCharacter++: Towards Comprehensive Benchmark for Realistic Role-Playing Agents" 🔥
-
🔥🔥🔥 This repository curates research on Weak-to-Strong Generalization across LLMs, multimodal learning, and beyond, focusing on how strong models learn from weak supervision and surpass their teach…
-
replication-data-and-code-when-LLMs-reliable-empathic-communication Public
Forked from aakriti1kumar/replication-data-and-code-when-LLMs-reliable-empathic-communicationReplication data and code for the paper: When LLMs are Reliable for Judging Empathic Communication
Jupyter Notebook MIT License UpdatedSep 30, 2025 -
OmniCharacter Public
[ACL25] Official codebase for "OmniCharacter:Towards Immersive Role-Playing Agents with Seamless Speech-Language Personality Interaction" 🔥
-
GLSCL Public
[TIP25] Code for "Text-Video Retrieval with Global-Local Semantic Consistent Learning"
-
UMP_TVR Public
[TCSVT24] The implementation of paper "UMP: Unified Modality-aware Prompt Tuning for Text-Video Retrieval".
-
S2-Transformer Public
[IJCAI 2022] Official Pytorch code for paper “S2 Transformer for Image Captioning”
-
SPT Public
[TCSVT23] Official code for "SPT: Spatial Pyramid Transformer for Image Captioning".
-
-
SNLC Public
[PR23] The implementation of the paper ''Learning Visual Question Answering on Controlled Semantic Noisy Labels''
-
PKOL Public
[TIP 2022] Official code of paper “Video Question Answering with Prior Knowledge and Object-sensitive Learning”
-
DAST Public
[MM23] Code for paper "Depth-Aware Sparse Transformer for Video-Language Learning"
-
-