Papers by Isabelle Tellier
Chunk Different Kind of Spoken Discourse: Challenges for Machine Learning (2020.lrec-1)
Copied to clipboard
| Challenge: | Existing chunkers for spoken data are based on a corpus composed of monologues and spontaneous talk in interaction. |
| Approach: | They propose to use CRFs to develop a chunker for spoken data . the chunker is based on a small corpus composed of two kinds of discourse . |
| Outcome: | The proposed chunker is based on a spoken corpus composed of monologue and spontaneous talk in interaction. |
ANCOR-AS: Enriching the ANCOR Corpus with Syntactic Annotations (L18-1)
Copied to clipboard
| Challenge: | ANCOR-AS is an enriched version of the ANCor corpus that adds syntactic annotations in addition to the existing coreference and speech transcription ones. |
| Approach: | They propose to use syntactic annotations in addition to existing coreference and speech transcription annotations to improve detection of mentions. |
| Outcome: | The proposed version adds syntactic annotations to existing coreference and speech transcription annotations and is released in a new TEI-compliant XML format. |