Papers by Abel Salinas
The Butterfly Effect of Altering Prompts: How Small Changes and Jailbreaks Affect Large Language Model Performance (2024.findings-acl)
Copied to clipboard
| Challenge: | Large Language Models (LLMs) are used to label data across many domains and for myriad tasks. |
| Approach: | They ask large language models to label data using a series of decisions by practitioners . they find that even the smallest perturbations can change the LLM's answer . |
| Outcome: | The proposed model can be used to quickly get a response for arbitrary tasks. |