Papers by Ron Yosef
EditInspector: A Benchmark for Evaluation of Text-Guided Image Edits (2025.acl-long)
Copied to clipboard
| Challenge: | Text-guided image editing is becoming increasingly widespread . current models struggle to evaluate edits comprehensively and often hallucinate when describing changes. |
| Approach: | They propose a novel framework to evaluate edits based on human annotations . they use a template to collect human annotation data and validate the results . |
| Outcome: | The proposed methods outperform current models in artifact detection and difference caption generation. |
ParallelPARC: A Scalable Pipeline for Generating Natural-Language Analogies (2024.naacl-long)
Copied to clipboard
| Challenge: | Analogy-making is a central to human cognition, allowing us to abstract information and understand novel situations in terms of familiar ones. |
| Approach: | They propose a pipeline to generate paragraph-based analogies using large language models and large language distractors. |
| Outcome: | The proposed pipeline outperforms existing models in binary and multiple-choice settings and shows that humans outperformed the best models after a light supervision. |
IRFL: Image Recognition of Figurative Language (2023.findings-emnlp)
Copied to clipboard
| Challenge: | Figures of speech are ubiquitous in many forms of discourse, allowing people to convey complex, abstract ideas and evoke emotion. |
| Approach: | They develop a dataset for multimodal figurative language understanding using human annotation and an automatic pipeline to generate a multimodal dataset. |
| Outcome: | The proposed dataset performs better than human vision and language models compared with a human dataset . |