A Diachronic Corpus for Literary Style Analysis (L18-1)

Copied to clipboard

Challenge: Temporal style analysis is not widely taken into account, says aaron daelemans . he says it is important to consider the possibility of an author's style frequently changing over time . daelemens: synchronic style analysis requires accurate time-stamped data .
Approach: They propose a resource for diachronic style analysis in particular the analysis of literary authors over time.
Outcome: The proposed resource can be used to analyze literary authors over time.

Similar Papers

DRIFT: A Toolkit for Diachronic Analysis of Scientific Literature (2021.emnlp-demo)

Copied to clipboard

Challenge: a tool for the diachronic analysis of research corpora is open source . the tool is based on well-cited research, with some of our own methods added for good measure.
Approach: They propose an application for the diachronic analysis of research corpora . they collate key words from well-cited research and add some of their own methods .
Outcome: The proposed tool is open source and easy to use.
Diachronic word embeddings and semantic shifts: a survey (C18-1)

Copied to clipboard

Challenge: Existing methods for tracing time-related semantic shifts with word embedding models lack the cohesion, common terminology and shared practices of more established areas of natural language processing.
Approach: They propose several axes along which these methods can be compared and propose a framework for comparison.
Outcome: The proposed methods are compared with existing methods and outline their main challenges and potential applications.
The EDGeS Diachronic Bible Corpus (2020.lrec-1)

Copied to clipboard

Challenge: EDGeS is a diachronic and parallel corpus of Bible translations in Dutch, English, German and Swedish . it is intended to be used for longitudinal studies of complex verb constructions in Germanic .
Approach: They present the EDGeS Diachronic Bible Corpus, a diachronic corpus of Bible translations in Dutch, English, German and Swedish . they use a synchronically and synchronly parallel corpus to study complex verb constructions in Germanic .
Outcome: The EDGeS is a diachronic and parallel corpus of Bible translations in Dutch, English, German and Swedish spanning six and a half centuries.
Possessors Change Over Time: A Case Study with Artworks (D18-1)

Copied to clipboard

Challenge: Existing methods to extract possession relations from Wikipedia articles can be used to extract possessors over time.
Approach: They propose to extract possession relations from Wikipedia articles and temporal information indicating when these relations are true.
Outcome: The proposed annotation scheme yields many possessors over time for a given artwork, and an LSTM ensemble can automate the task.
The DReaM Corpus: A Multilingual Annotated Corpus of Grammars for the World’s Languages (2020.lrec-1)

Copied to clipboard

Challenge: Until recently, language descriptions were available in paper form only, with indexes as the only search aid.
Approach: They propose to digitize a multilingual corpus of language descriptions and annotate it with various meta, word, and text attributes to make searching and analysis easier and more useful.
Outcome: The proposed corpus is searchable through a couple of well-established corpus infrastructures.
Neural Temporality Adaptation for Document Classification: Diachronic Word Embeddings and Domain Adaptation Models (P19-1)

Copied to clipboard

Challenge: Recent studies show that document classifiers can become more stable over time when trained in ways that account for temporal variations.
Approach: They propose a method for embedding diachronic word embedds into document classification models . they propose 'time-driven neural classification model' that accounts for temporal variations .
Outcome: The proposed model can be trained on six corpora and make it more robust over time.
Towards Actual (Not Operational) Textual Style Transfer Auto-Evaluation (D19-55)

Copied to clipboard

Challenge: elucidates the dangerous current state of style transfer auto-evaluation research.
Approach: They propose ways to aggregate the three metrics into one evaluator.
Outcome: The proposed method could be used to aggregate the three metrics into one evaluator.
Detecting, Generating, and Evaluating in the Writing Style of Different Authors (2025.naacl-srw)

Copied to clipboard

Challenge: In recent years, stylometry has been investigated in many different fields.
Approach: They propose to use sentences from different books to generate and evaluate stylistic texts according to the authors' writing styles.
Outcome: The proposed model can detect, generate, and evaluate documents according to the authors' writing styles with unpaired samples.
Analyzing Continuous Semantic Shifts with Diachronic Word Similarity Matrices (2025.coling-main)

Copied to clipboard

Challenge: Existing methods to analyze word sense proportions are insufficient for understanding semantic shifts . et al., 2018: semantic shift and its effects.
Approach: They propose a framework for how semantic shifts occur over multiple time periods by using word embeddings.
Outcome: The proposed framework can analyze semantic shifts over multiple time periods using word embeddings.
A Methodology for Building a Diachronic Dataset of Semantic Shifts and its Application to QC-FR-Diac-V1.0, a Free Reference for French (2022.lrec-1)

Copied to clipboard

Challenge: Existing algorithms to detect semantic shifts have been criticized for their difficulty in evaluating them.
Approach: They propose a method for building a reference dataset for semantic shift detection . they use a word-sense disambiguation model to associate a date of first appearance to all senses of a term .
Outcome: The proposed method is based on a word-sense disambiguation model . significant changes in sense distributions and stability are detected . the resulting words are inspected by experts using a dedicated interface .

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations