Challenge: This study examines the literary perceptions of noise during the Scandinavian "Modern Breakthrough" period (1870-1899).
Approach: They propose a framework for detecting and categorizing noise in literary texts from the late 19th century.
Outcome: The proposed framework can be applied to Danish and Norwegian literature from the late 19th century.

Similar Papers

Fact from Fiction: Finding Serialized Novels in Newspapers (2025.acl-srw)

Copied to clipboard

Challenge: Among underrepresented but widely read forms are serialized fiction and feuilleton novels embedded in newspapers rather than published as standalone volumes.
Approach: They propose to annotate 1,394 articles and evaluate classification pipelines using both selected linguistic features and embeddings to identify serialized fiction and feuilleton fiction.
Outcome: The proposed methods achieve F1-scores of 0.91 in an annotated dataset of 1,394 articles and support the construction of alternative literary corpora and contribute to work on modeling the fiction–nonfiction boundary at scale.
Literary Event Detection (P19-1)

Copied to clipboard

Challenge: a new dataset of literary events is presented to examine the nature of narratives . literature presents a number of challenges for existing systems, including complex narration .
Approach: They propose a dataset of literary events that are depicted as taking place within the imagined space of a novel.
Outcome: The proposed model achieves an F1 score of 73.9 for prestige and popularity . the best performing model achieve a score of 79.9 for prestige compared to the previous model .
Letters From the Past: Modeling Historical Sound Change Through Diachronic Character Embeddings (2022.acl-long)

Copied to clipboard

Challenge: a great deal of work has been done on NLP approaches to lexical semantic change detection, but other aspects of language change have received less attention from the NLP community.
Approach: They propose to compare the relative distance through time between the distributions of the characters involved before and after a sound change has taken place.
Outcome: The proposed method can trace the well-known historical change of lenition of plosives in Danish historical sources and identify several of the changes under consideration and uncover meaningful contexts in which they appeared.
What time is it? Temporal Analysis of Novels (2020.emnlp-main)

Copied to clipboard

Challenge: a novel based on the flow of time provides a framework for understanding the text . a computational approach to annotate a book's lines with wall clock times is needed to understand the flow through time.
Approach: They propose to annotate each line of a book with wall clock times . they use a data set of hourly time phrases from 52,183 fictional books .
Outcome: The proposed method improves upon baselines by over two hours and can partition a book into segments that correspond to a particular time-of-day.
Development and Evaluation of Pre-trained Language Models for Historical Danish and Norwegian Literary Texts (2024.lrec-main)

Copied to clipboard

Challenge: et al., 2019) develop and evaluate the first pre-trained language models specifically tailored for historical Danish and Norwegian texts.
Approach: They develop and evaluate pre-trained language models specifically tailored for historical Danish and Norwegian texts.
Outcome: The proposed model outperforms models trained on historical Danish and Norwegian literature in two downstream NLP tasks.
CLAUSE-ATLAS: A Corpus of Narrative Information to Scale up Computational Literary Analysis (2024.lrec-main)

Copied to clipboard

Challenge: XIX and XX century English novels annotated automatically contain 41,715 labeled clauses . a new approach to analyze novels based on clauses captures structural patterns within books, as well as qualitative differences between them.
Approach: They propose to use a corpus of XIX and XX century English novels annotated automatically to study stories as sequences of eventive, subjective and contextual information.
Outcome: The proposed method captures structural patterns within books, as well as qualitative differences between them.
Measuring Information Propagation in Literary Social Networks (2020.emnlp-main)

Copied to clipboard

Challenge: a gap in computational work to support the "Miss Havisham is dead" "She died" research focuses on the representation of social networks in literature .
Approach: They propose a pipeline for measuring information propagation in literature . they analyze the dynamics of information propagations in over 5,000 works of fiction .
Outcome: The proposed pipeline analyzes the dynamics of information propagation in 5,000 works of fiction and finds that women fill structural holes connecting different communities more frequently than men.
Assessing the State of the Art in Scene Segmentation (2025.naacl-long)

Copied to clipboard

Challenge: Recent advances in scene segmentation have made it difficult to detect scenes in literary texts.
Approach: They propose to modify existing models to improve detection of scenes in literary texts . they propose to use a training sample generation scheme to alleviate this problem .
Outcome: The proposed model is more robust to different types of texts, while its overall performance is slightly worse than that of BERT-based models.
Towards a music-language mapping (L18-1)

Copied to clipboard

Challenge: a novel research idea investigates the possibility of musical input to speech interaction systems.
Approach: They propose a musical language processing idea that investigates the possibility of musical input to speech interaction systems.
Outcome: The proposed method could be used to map musical pieces and dialogues based on frequency of musical patterns . the proposed method is universal among different languages and easy to learn for musicians .
Dying or Departing? Euphemism Detection for Death Discourse in Historical Texts (2025.coling-main)

Copied to clipboard

Challenge: euphemisms are a linguistic device used to soften discussions of uncomfortable topics . euphorias are used to refer to death in a less direct manner during a period of secularization .
Approach: They propose to use a corpus of Danish and Norwegian novels to detect death-related euphemisms . they use pre-trained language models to detect euphoric and literal references to death .
Outcome: The proposed method improves on state-of-the-art language models.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations