Papers by Gyeongmin Kim
KoCHET: A Korean Cultural Heritage Corpus for Entity-related Tasks (2022.coling-1)
Copied to clipboard
| Challenge: | Existing corpus for entity-related tasks is limited in terms of application and cannot be used for entity recognition. |
| Approach: | They propose to use a Korean cultural heritage corpus for the typical entity-related tasks named entity recognition (NER), relation extraction (RE) and entity typing (ET) . |
| Outcome: | The proposed corpus makes it more useful in terms of cultural heritage and provides practical insights in terms linguistic analysis. |
QUAK: A Synthetic Quality Estimation Dataset for Korean-English Neural Machine Translation (2022.coling-1)
Copied to clipboard
| Challenge: | despite its high utility, there are limitations concerning manual QE data creation. |
| Approach: | They propose to generate a Korean-English QE dataset that is fully automatic . they find that the algorithm is more accurate and faster than manual QE . |
| Outcome: | The proposed datasets show that they scale up to 1.58M and 6.58M, respectively, and show that the results are significantly better when compared to the previous datasets. |