PRADALab_KAUST
Popular repositories Loading
-
-
repeat-curse-llm
repeat-curse-llm Public[ACL 2025 Findings] Understanding the Repeat Curse in Large Language Models from a Feature Perspective
-
LLM-Persona-Steering
LLM-Persona-Steering PublicOfficial code of "Exploring the Personality Traits of LLMs through Latent Features Steering"
-
LLM-sycophancy
LLM-sycophancy Public[AAAI'26 Main🎉] Official code of "When Truth Is Overridden: Uncovering the Internal Origins of Sycophancy in Large Language Models"
Repositories
- continuous-adv-ICL Public Forked from fshp971/continuous-adv-ICL
[ICLR 2026] Official repository for "Understanding and Improving Continuous LLM Adversarial Training via In-context Learning Theory"
- adv-ICL Public Forked from fshp971/adv-ICL
[NeurIPS 2025] Official repository for "Short-length Adversarial Training Helps LLMs Defend Long-length Jailbreak Attacks: Theoretical and Empirical Evidence"
- LLM-sycophancy Public
[AAAI'26 Main🎉] Official code of "When Truth Is Overridden: Uncovering the Internal Origins of Sycophancy in Large Language Models"
- flashdp Public
- Fraud-R1 Public
[ACL 2025 Findings] Fraud-R1 : A Multi-Round Benchmark for Assessing the Robustness of LLM Against Augmented Fraud and Phishing Inducements
- repeat-curse-llm Public
[ACL 2025 Findings] Understanding the Repeat Curse in Large Language Models from a Feature Perspective
- ECBM Public
- zo2 Public Forked from liangyuwang/zo2
ZO2 (Zeroth-Order Offloading): Full Parameter Fine-Tuning 175B LLMs with 18GB GPU Memory
- draft Public Forked from kaustpradalab/zo2
Privately Fine-Tuning Extremely Large Language Models with Zeroth-Order Offloading
- vanilla-RLAIF-pipeline Public Forked from mengdi-li/vanilla-RLAIF-pipeline
An implementation of a vanilla RLAIF pipeline, utilizing GPT-2-Large for the summarization task with the TL;DR dataset.
Top languages
Loading…
Most used topics
Loading…