Joe Barrow
发表论文 32 篇 · 总被引 2072 次 · h-index 15
代表论文
- Bias and Fairness in Large Language Models: A Survey (2023 · Computational Linguistics · 被引 1291)
- Evaluation Examples are not Equally Informative: How should that change NLP Leaderboards? (2021 · Annual Meeting of the Association for Computational Linguistics · 被引 153)
- Personalization of Large Language Models: A Survey (2024 · arXiv.org · 被引 134)
- AutoDAN: Interpretable Gradient-Based Adversarial Attacks on Large Language Models (2023 · 被引 124)
- AutoDAN: Automatic and Interpretable Adversarial Attacks on Large Language Models (2023 · arXiv.org · 被引 118)
- Self-Debiasing Large Language Models: Zero-Shot Recognition and Reduction of Stereotypes (2024 · arXiv.org · 被引 80)