PRADALab_KAUST
Popular repositories Loading
-
-
repeat-curse-llm
repeat-curse-llm Public[ACL 2025 Findings] Understanding the Repeat Curse in Large Language Models from a Feature Perspective
-
LLM-Persona-Steering
LLM-Persona-Steering PublicOfficial code of "Exploring the Personality Traits of LLMs through Latent Features Steering"
-
LLM-sycophancy
LLM-sycophancy Public[AAAI'26 Main🎉] Official code of "When Truth Is Overridden: Uncovering the Internal Origins of Sycophancy in Large Language Models"
Repositories
- understanding-aha-moments Public
Research code for Understanding Aha Moments: From External Observations to Internal Mechanisms
- continuous-adv-ICL Public Forked from fshp971/continuous-adv-ICL
[ICLR 2026] Official repository for "Understanding and Improving Continuous LLM Adversarial Training via In-context Learning Theory"
- LLM-sycophancy Public
[AAAI'26 Main🎉] Official code of "When Truth Is Overridden: Uncovering the Internal Origins of Sycophancy in Large Language Models"
- repeat-curse-llm Public
[ACL 2025 Findings] Understanding the Repeat Curse in Large Language Models from a Feature Perspective
Most used topics
Loading…