Skip to content
View davidwang200099's full-sized avatar
  • Shanghai Jiao Tong University
  • 800 Dongchuan Road, Shanghai, China

Highlights

  • Pro

Organizations

@SJTU-CSE

Block or report davidwang200099

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse

Starred repositories

Showing results

A Minimal and Elegant Framework & Tutorial for Real-Time Interactive World Models

Python 738 19 Updated Jun 15, 2026

Mechanistic Interpretability toolkit for Vision-Language-Action models

Python 20 4 Updated Jul 8, 2026

From Human Videos to Robot Manipulation: A Survey on Scalable Vision-Language-Action Learning with Human-Centric Data

15 Updated Jun 2, 2026
Python 805 47 Updated Jul 3, 2026

🍑 relsim: Relational Visual Similarity | pip install relsim 🌍 (CVPR 2026)

Python 87 2 Updated Jul 22, 2026

VLA model interpretability tools

Python 175 8 Updated Mar 30, 2026

中国科研常用LaTeX模板集

Python 2,553 246 Updated Jul 20, 2026

启智平台任务管理 CLI:资源查询、任务提交、日志查看和 MCP/agent workflow

Python 109 28 Updated Jul 17, 2026

Official Repo of The Great March Project. https://www.rhos.ai/research/gm-100

21 1 Updated Jun 17, 2026

[AAAI 2026] The Official Implementation for "Anomagic: Crossmodal Prompt-driven Zero-shot Anomaly Generation"

Python 157 8 Updated Apr 25, 2026

[CSUR] A Survey on Video Diffusion Models

2,304 116 Updated Jun 22, 2026

[ICCV 2025] Official implementation of "Holistic Unlearning Benchmark: A Multi-Faceted Evaluation for Text-to-Image Diffusion Model Unlearning"

Python 23 1 Updated Jan 2, 2026

A comprehensive reading list for Emotion Recognition in Conversations

282 44 Updated Feb 6, 2024

awesome papers in LLM interpretability

624 21 Updated Aug 20, 2025

awesome SAE papers

79 2 Updated May 24, 2025

[TMLR 2025🔥] A survey for the autoregressive models in vision.

804 23 Updated May 5, 2026

collection of diffusion model papers categorized by their subareas

2,218 102 Updated Mar 16, 2026

😎 up-to-date & curated list of awesome 3D Visual Grounding papers, methods & resources.

283 6 Updated Jan 14, 2026

EO: Open-source Unified Embodied Foundation Model Series

Jupyter Notebook 291 29 Updated Nov 12, 2025

Microsoft BASIC for 6502 Microprocessor - Version 1.1

Assembly 4,497 507 Updated Sep 3, 2025

A resource repository for machine unlearning in large language models

617 32 Updated Jul 23, 2026

A curated collection of resources focused on the Mechanistic Interpretability (MI) of Large Multimodal Models (LMMs). This repository aggregates surveys, blog posts, and research papers that explor…

215 5 Updated Mar 4, 2026

CLIP (Contrastive Language-Image Pretraining), Predict the most relevant text snippet given an image

Jupyter Notebook 34,059 4,041 Updated Mar 25, 2026

Advanced AI Explainability for computer vision. Support for CNNs, Vision Transformers, Classification, Object detection, Segmentation, Image similarity and more.

Python 12,925 1,705 Updated Jul 10, 2026

Code for "Scaling Language-Free Visual Representation Learning" paper (Web-SSL).

Python 214 12 Updated Mar 20, 2026

Simple RL training for reasoning

Python 3,870 285 Updated Dec 23, 2025

A curated list of state-of-the-art research in embodied AI, focusing on vision-language-action (VLA) models, vision-language navigation (VLN), and related multimodal learning approaches.

3,400 160 Updated Jul 7, 2026

Code for Scaling Language-Free Visual Representation Learning (WebSSL)

244 2 Updated Apr 24, 2025

An aggregation of human motion understanding research.

288 20 Updated Jul 24, 2026
Next