Papers by Lluis Marquez
Diable: Efficient Dialogue State Tracking as Operations on Tables (2023.findings-acl)
Copied to clipboard
| Challenge: | Existing systems for dialogue state tracking use the full dialogue history as input and generate the entire state from scratch at each dialogue turn. |
| Approach: | They propose a task formalisation that represents the dialogue state as a table and formalises it as 'table manipulation task' they represent the dialogue as if it were a list with all the slots and generate the entire state from scratch at each dialogue turn. |
| Outcome: | The proposed system outperforms existing systems while maintaining competitive accuracy. |
Unraveling and Mitigating Safety Alignment Degradation of Vision-Language Models (2025.findings-acl)
Copied to clipboard
Qin Liu, Chao Shang, Ling Liu, Nikolaos Pappas, Jie Ma, Neha Anna John, Srikanth Doss, Lluis Marquez, Miguel Ballesteros, Yassine Benajiba
| Challenge: | LLaVA-7B demonstrated a decline in safety alignment ability on multi-modal inputs compared to its LLM backbone. |
| Approach: | They propose a method to recover alignment ability from LLM backbone while preserving functional capabilities of VLMs. |
| Outcome: | The proposed framework recovers alignment ability that is inherent in the LLM backbone with minimal impact on fluency and linguistic capabilities of pre-trained VLMs. |
Understanding and Improving Information Preservation in Prompt Compression for LLMs (2025.findings-emnlp)
Copied to clipboard
| Challenge: | Recent advances in large language models have enabled their successful application to a broad range of tasks. |
| Approach: | They propose a framework that allows for in-depth analysis of prompt compression methods. |
| Outcome: | The proposed framework analyzes state-of-the-art soft and hard compression methods . it shows that some fail to preserve key details from the original prompt, limiting performance on complex tasks. |
Factual Confidence of LLMs: on Reliability and Robustness of Current Estimators (2024.acl-long)
Copied to clipboard
| Challenge: | Large Language Models (LLMs) tend to be unreliable on fact-based answers. |
| Approach: | They propose a framework for comparing LLMs' confidence over fact-based answers with hidden-state probes that are more reliable than hidden-status probes. |
| Outcome: | The proposed methods show that hidden-state probes provide the most reliable confidence estimates despite requiring access to weights and supervision data. |