Papers by Qijun Tan

3 papers
High Quality Rather than High Model Probability: Minimum Bayes Risk Decoding with Neural Metrics (2022.tacl-1)

Copied to clipboard

Challenge: Neural machine translations are ranked below human translations in professional evaluations .
Approach: They apply minimum bayes risk decoding to optimize different metrics of translation quality . they show that model estimates and translation quality only vaguely correlate .
Outcome: The proposed method improves human translations with different models and metric.
Experts, Errors, and Context: A Large-Scale Study of Human Evaluation for Machine Translation (2021.tacl-1)

Copied to clipboard

Challenge: a large study of machine translation systems shows poor evaluation procedures can lead to erroneous conclusions.
Approach: They propose an evaluation methodology grounded in explicit error analysis based on the Multidimensional Quality Metrics framework.
Outcome: The proposed evaluation methodology outperforms crowd workers in two languages . it shows that human-based metrics outperformed crowd workers .
Improving Diversity of Demographic Representation in Large Language Models via Collective-Critiques and Self-Voting (2023.emnlp-main)

Copied to clipboard

Challenge: Existing studies on diversity in large language models focus on the understudied class of fairness and inclusion concern in LLMs.
Approach: They propose a technique to measure diversity in generated responses along people and culture axes by collective-critique and self-voting.
Outcome: The proposed approach outperforms baseline methods and human evaluations with human and automated evaluations.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations