Papers by Bruno Cuconato

    1 papers
    Text Mining for History: first steps on building a large dataset (L18-1)

    Copied to clipboard

    Challenge: a new corpus on the history domain is being created to mine text in the domain . primary motivation for the project is the need to query the material in a non-linear way .
    Approach: They propose to use a Brazilian historical-biographical dictionary as a resource for text mining.
    Outcome: The proposed corpus is a reference work on the Brazilian history domain . it contains almost 12 millions tokens in about three hundred thousand sentences . the authors argue that the proposed corpu is linguistically motivated .

    What is GenGO?

    GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

    Information

    About
    Limitations