Papers by Joachim Daiber

2 papers
MKQA: A Linguistically Diverse Benchmark for Multilingual Open Domain Question Answering (2021.tacl-1)

Copied to clipboard

Challenge: Existing multilingual QA datasets lack linguistic diversity and comparable evaluation between languages.
Approach: They propose a multilingual question-answer evaluation set with 10k English queries and human translations of them into 25 additional languages and dialects.
Outcome: The proposed model is based on a multilingual knowledge questions and answers evaluation set with 26 languages.
DispatchQA: A Benchmark for Small Function Calling Language Models in E-Commerce Applications (2025.emnlp-industry)

Copied to clipboard

Challenge: DispatchQA is a benchmark to evaluate how well small language models (SLMs) translate openended search queries into executable API calls via explicit function calling.
Approach: They propose a benchmark to evaluate how well small language models translate openended search queries into executable API calls via explicit function calling.
Outcome: The proposed benchmark aims to evaluate how well small language models (SLMs) translate openended search queries into executable API calls via explicit function calling.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations