Papers by Joy Mahapatra
sudoLLM: On Multi-role Alignment of Language Models (2025.findings-emnlp)
Copied to clipboard
| Challenge: | a framework that allows users to control access rights has not been extensively studied in the large language model realm. |
| Approach: | They propose a framework that allows users to control access rights in a multi-role manner. |
| Outcome: | The proposed framework improves alignment, generalization and resistance to prefix-based jailbreaking attacks. |