PromptBench: Towards Evaluating the Robustness of Large Language Models on Adversarial Prompts Paper • 2306.04528 • Published Jun 7, 2023 • 3
Improving Generalization of Adversarial Training via Robust Critical Fine-Tuning Paper • 2308.02533 • Published Aug 1, 2023
Large Language Models Understand and Can be Enhanced by Emotional Stimuli Paper • 2307.11760 • Published Jul 14, 2023 • 1
DyVal: Dynamic Evaluation of Large Language Models for Reasoning Tasks Paper • 2309.17167 • Published Sep 29, 2023 • 1
PromptBench: A Unified Library for Evaluation of Large Language Models Paper • 2312.07910 • Published Dec 13, 2023 • 16