Skip to content
@kaustpradalab

PRADALab_KAUST

Provable Responsible AI and Data Analytics (PRADA) Lab, KAUST

Popular repositories Loading

  1. research-handboook research-handboook Public

    82 4

  2. Fraud-R1 Fraud-R1 Public

    [ACL 2025 Findings] Fraud-R1 : A Multi-Round Benchmark for Assessing the Robustness of LLM Against Augmented Fraud and Phishing Inducements

    Python 37 3

  3. repeat-curse-llm repeat-curse-llm Public

    [ACL 2025 Findings] Understanding the Repeat Curse in Large Language Models from a Feature Perspective

    Python 22 2

  4. LLM-Persona-Steering LLM-Persona-Steering Public

    Official code of "Exploring the Personality Traits of LLMs through Latent Features Steering"

    Python 19 3

  5. LLM-sycophancy LLM-sycophancy Public

    [AAAI'26 Main🎉] Official code of "When Truth Is Overridden: Uncovering the Internal Origins of Sycophancy in Large Language Models"

    Python 12 2

  6. flashdp flashdp Public

    Python 8 3

Repositories

Showing 10 of 24 repositories
  • understanding-aha-moments Public

    Research code for Understanding Aha Moments: From External Observations to Internal Mechanisms

    kaustpradalab/understanding-aha-moments's past year of commit activity
    Python 0 0 0 0 Updated Sep 13, 2026
  • continuous-adv-ICL Public Forked from fshp971/continuous-adv-ICL

    [ICLR 2026] Official repository for "Understanding and Improving Continuous LLM Adversarial Training via In-context Learning Theory"

    kaustpradalab/continuous-adv-ICL's past year of commit activity
    Python 0 1 0 0 Updated Apr 14, 2026
  • adv-ICL Public Forked from fshp971/adv-ICL

    [NeurIPS 2025] Official repository for "Short-length Adversarial Training Helps LLMs Defend Long-length Jailbreak Attacks: Theoretical and Empirical Evidence"

    kaustpradalab/adv-ICL's past year of commit activity
    Python 0 1 0 0 Updated Feb 19, 2026
  • LLM-sycophancy Public

    [AAAI'26 Main🎉] Official code of "When Truth Is Overridden: Uncovering the Internal Origins of Sycophancy in Large Language Models"

    kaustpradalab/LLM-sycophancy's past year of commit activity
    Python 12 2 1 0 Updated Nov 11, 2025
  • flashdp Public
    kaustpradalab/flashdp's past year of commit activity
    Python 8 Apache-2.0 3 0 0 Updated Jul 1, 2025
  • Fraud-R1 Public

    [ACL 2025 Findings] Fraud-R1 : A Multi-Round Benchmark for Assessing the Robustness of LLM Against Augmented Fraud and Phishing Inducements

    kaustpradalab/Fraud-R1's past year of commit activity
    Python 37 3 1 0 Updated Jun 29, 2025
  • repeat-curse-llm Public

    [ACL 2025 Findings] Understanding the Repeat Curse in Large Language Models from a Feature Perspective

    kaustpradalab/repeat-curse-llm's past year of commit activity
    Python 22 2 0 0 Updated Jun 13, 2025
  • ECBM Public
    kaustpradalab/ECBM's past year of commit activity
    Python 2 0 1 0 Updated May 27, 2025
  • zo2 Public Forked from liangyuwang/zo2

    ZO2 (Zeroth-Order Offloading): Full Parameter Fine-Tuning 175B LLMs with 18GB GPU Memory

    kaustpradalab/zo2's past year of commit activity
    Python 3 Apache-2.0 18 0 0 Updated Apr 13, 2025
  • draft Public Forked from kaustpradalab/zo2

    Privately Fine-Tuning Extremely Large Language Models with Zeroth-Order Offloading

    kaustpradalab/draft's past year of commit activity
    Python 0 Apache-2.0 18 0 0 Updated Mar 10, 2025

Most used topics

Loading…