Papers by Anna Squicciarini
The Task Shield: Enforcing Task Alignment to Defend Against Indirect Prompt Injection in LLM Agents (2025.acl-long)
Copied to clipboard
| Challenge: | Large Language Model (LLM) agents are becoming conversational assistants . indirect prompt injection attacks pose a critical threat to these systems . |
| Approach: | They propose a novel and orthogonal perspective that reframes agent security . they propose 'task shield' that verifies whether each instruction and tool call contributes to user objectives . |
| Outcome: | The proposed defense reduces attack success rates while maintaining high task utility on the AgentDojo benchmark. |
A Semantics-based Approach to Disclosure Classification in User-Generated Online Content (2020.findings-emnlp)
Copied to clipboard
| Challenge: | Existing algorithms for self-disclosure identification and classification are challenging due to the relative anonymity of social networking sites and lack of non-verbal cues to signal thoughts or feelings. |
| Approach: | They propose an approach to detect emotional and informational self-disclosure in natural language by using frame semantics to identify lexical units and their semantic roles. |
| Outcome: | The proposed method improves on reddit data and provides insights into the drivers of disclosure behaviors. |