Highlights
- Pro
Stars
The source code for the paper "ModelLens: Finding the Best for Your Task from Myriads of Models"
FabScore: Fine-Grained Evaluation of Fabrications in Automated AI Research
A web app for ranking computer science departments according to their research output in selective venues, and for finding active faculty across a wide range of areas.
[ICLR 2024] The official implementation of our ICLR2024 paper "AutoDAN: Generating Stealthy Jailbreak Prompts on Aligned Large Language Models".
[ICLR 2025 Spotlight] The official implementation of our ICLR2025 paper "AutoDAN-Turbo: A Lifelong Agent for Strategy Self-Exploration to Jailbreak LLMs".
Code for paper "Unraveling Cross-Modality Knowledge Conflicts in Large Vision-Language Models."
This repository is the official implementation of "Red Teaming Language Models for Processing Contradictory Dialogues"
[EMNLP 2025 Findings] Familiarity-aware Evidence Compression for Retrieval Augmented Generation
Code and models for EMNLP 2024 paper "WPO: Enhancing RLHF with Weighted Preference Optimization"
[EMNLP 2024] mDPO: Conditional Preference Optimization for Multimodal Large Language Models.
World Model based Autonomous Driving Platform in CARLA 🚗
Code for our NAACL 2024 Paper "Cognitive Overload: Jailbreaking Large Language Models with Overloaded Logical Thinking"
[EMNLP 2023] Bridging Continuous and Discrete Spaces: Interpretable Sentence Representation Learning via Compositional Operations
[EMNLP 2023] VIPHY: Probing “Visible” Physical Commonsense Knowledge
[ACL 2023] Robust Natural Language Understanding with Residual Attention Debiasing
Official Implementation for "Can NLI Provide Proper Indirect Supervision for Low-resource Biomedical Relation Extraction?" ACL 2023
Multi-hop Evidence Retrieval for Cross-document Relation Extraction