Papers by Daniel Loebenberger
Detection of Adversarial Prompts with Model Predictive Entropy (2026.findings-eacl)
Copied to clipboard
| Challenge: | Large Language Models (LLMs) are increasingly deployed in high-impact scenarios raising concerns about their safety and security. |
| Approach: | They propose an attack-agnostic pipeline for detecting adversarial inputs without prior knowledge of attack specifications. |
| Outcome: | The proposed pipeline outperforms traditional defenses in terms of adaptability and resource efficiency. |