Papers by Wilbert Heeringa

2 papers
PoS Tagging, Lemmatization and Dependency Parsing of West Frisian (2022.lrec-1)

Copied to clipboard

Challenge: a lemmatizer/PoS tagger/dependency parser for west frisian is released as a web app and as . web service.
Approach: They propose a lemmatizer/PoS tagger/dependency parser for West Frisian using a corpus of 44,714 words in 3,126 sentences that were annotated according to the guidelines of Universal Dependencies version 2.
Outcome: The proposed lemmatizer/PoS tagger/dependency parser performs better than the previous version of Oersetter . the current corpus contains 44,714 words in 3,126 sentences .
The Boarnsterhim Corpus: A Bilingual Frisian-Dutch Panel and Trend Study (L18-1)

Copied to clipboard

Challenge: a corpus of 250 hours of speech in both west frisian and Dutch is being developed . the corpus is a sociolinguistic corpus based on the recordings of four generations of bilingual speakers .
Approach: This paper describes the Boarnsterhim Corpus project which started in 2016 . it aims to make available 250 hours of speech in both west frisian and Dutch by same speakers .
Outcome: The Boarnsterhim Corpus is a sociolinguistic corpus of west frisian and Dutch speakers . it spans four generations and includes panel and trend data .

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations