Papers by Zidi Xiu
DELPHI: Data for Evaluating LLMs’ Performance in Handling Controversial Issues (2023.emnlp-industry)
Copied to clipboard
| Challenge: | a recent study of controversy-handling in large language models (LLMs) has shown that people may become increasingly dependent on such systems for information. |
| Approach: | They propose to construct a controversial questions dataset using a subset of a publicly available dataset. |
| Outcome: | The proposed dataset presents challenges concerning knowledge recency, safety, fairness, and bias. |