Papers by Konrad Wojtasik

2 papers
BEIR-PL: Zero Shot Information Retrieval Benchmark for the Polish Language (2024.lrec-main)

Copied to clipboard

Challenge: Existing multilingual evaluation benchmarks focus on IR in the Polish language, but the Polish is a relatively new field due to the limited availability of Polish datasets.
Approach: They propose to establish large-scale resources for IR in the Polish language and translate them into a new benchmark which includes 13 datasets.
Outcome: The proposed benchmarks are based on 13 open IR datasets in Polish and are a pioneering development in this area.
Developing PUGG for Polish: A Modern Approach to KBQA, MRC, and IR Dataset Construction (2024.findings-acl)

Copied to clipboard

Challenge: Existing KBQA datasets are outdated and inefficient in human labor, and assisting tools like Large Language Models (LLM) are not utilized to reduce the workload.
Approach: They propose a semi-automated question answering task that uses structured knowledge graphs to answer extensive knowledge-intensive questions.
Outcome: The proposed approach includes KBQA, MRC, and Information Retrieval tasks for low-resource languages.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations