Papers by Adam Nohejl

4 papers
WorldCuisines: A Massive-Scale Benchmark for Multilingual and Multicultural Visual Question Answering on Global Cuisines (2025.naacl-long)

Copied to clipboard

Challenge: Vision Language Models struggle with cultural-specific knowledge, especially in languages other than English and in underrepresented cultural contexts.
Approach: They propose a visual question answering (VQA) dataset with text-image pairs across 30 languages and dialects and a training dataset.
Outcome: The proposed model performs better with correct location context, but struggles with adversarial contexts and predicting specific regional cuisines and languages.
Measuring the Robustness of Reference-Free Dialogue Evaluation Systems (2025.coling-main)

Copied to clipboard

Challenge: Advancements in dialogue systems powered by large language models have outpaced the development of reliable evaluation metrics.
Approach: They propose a benchmark to evaluate the robustness of reference-free dialogue metrics against four categories of adversarial attacks.
Outcome: The proposed benchmarks show that the two axes of reliability are not always aligned . the findings motivate the development of nuanced evaluation frameworks to address real-world dialogue challenges.
Beyond Film Subtitles: Is YouTube the Best Approximation of Spoken Vocabulary? (2025.coling-main)

Copied to clipboard

Challenge: Word frequency is a key variable in psycholinguistics, useful for modeling human familiarity with words . a recent study shows that frequency from YouTube subtitles is comparable to and often better than the best available resources.
Approach: They use YouTube subtitles to construct frequency norms for five languages . they find they are comparable to and often better than the best currently available resources .
Outcome: The proposed method improves on the best currently available resources for Chinese, English, Indonesian, Japanese, and Spanish.
CoAM: Corpus of All-Type Multiword Expressions (2025.acl-long)

Copied to clipboard

Challenge: Existing datasets for multiword expressions are inconsistently annotated, limited to a single type of MWE, or limited in size.
Approach: They propose to use a new interface to generate MWE annotations for the first time in a dataset of MWE identification.
Outcome: The proposed model outperforms existing models on the DiMSUM dataset.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations