Papers by Matthew Coole

2 papers
Infrastructure for Semantic Annotation in the Genomics Domain (2020.lrec-1)

Copied to clipboard

Challenge: a novel infrastructure for biomedical text mining combines NLP and corpus linguistics methods to provide a comprehensive corpus for literature-based discovery.
Approach: They propose a novel pipeline for the collection, annotation, storage, retrieval and analysis of biomedical and life sciences literature . it uses an updatable Gene Ontology Semantic Tagger and a NLP pipeline scheduler to collect and process the corpus.
Outcome: The proposed infrastructure allows for extreme-scale research on the open access PubMed Central archive.
LexiDB: Patterns & Methods for Corpus Linguistic Database Management (2020.lrec-1)

Copied to clipboard

Challenge: LexiDB is a tool for storing, managing and querying corpus data.
Approach: They propose to use LexiDB for storing, managing and querying corpus data.
Outcome: The proposed methods outperform existing tools for corpus queries and storage.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations