Papers by Isabelle Tellier

2 papers
Chunk Different Kind of Spoken Discourse: Challenges for Machine Learning (2020.lrec-1)

Copied to clipboard

Challenge: Existing chunkers for spoken data are based on a corpus composed of monologues and spontaneous talk in interaction.
Approach: They propose to use CRFs to develop a chunker for spoken data . the chunker is based on a small corpus composed of two kinds of discourse .
Outcome: The proposed chunker is based on a spoken corpus composed of monologue and spontaneous talk in interaction.
ANCOR-AS: Enriching the ANCOR Corpus with Syntactic Annotations (L18-1)

Copied to clipboard

Challenge: ANCOR-AS is an enriched version of the ANCor corpus that adds syntactic annotations in addition to the existing coreference and speech transcription ones.
Approach: They propose to use syntactic annotations in addition to existing coreference and speech transcription annotations to improve detection of mentions.
Outcome: The proposed version adds syntactic annotations to existing coreference and speech transcription annotations and is released in a new TEI-compliant XML format.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations