Analyzing Citation-Distance Networks for Evaluating Publication Impact (L18-1)

Copied to clipboard

Challenge: citation networks are used to study scholarly articles' semantic distances and their referencing patterns.
Approach: They propose to analyze the semantic distance of scholarly articles in a citation network to uncover patterns that reflect scientific impact.
Outcome: The proposed method combines semantic distance and content similarity to uncover scientific impact of articles in two different types of publications.

Similar Papers

Evaluating Scholarly Impact: Towards Content-Aware Bibliometrics (2021.emnlp-main)

Copied to clipboard

Challenge: Scientific, engineering, and technological (SET) innovations drive many positive advances in our modern economy, society, and life.
Approach: They propose a new metric that uses the content of the paper as a source of distant-supervision to quantify how much the cited-node informs the citing-n node.
Outcome: The proposed method achieves up to 103% improvement over the second-best method.
Beyond Citations: Corpus-based Methods for Detecting the Impact of Research Outcomes on Society (2020.lrec-1)

Copied to clipboard

Challenge: Existing methods for assessing the impact of research are ineffective for identifying impact beyond academia and text-based indicators beyond those that capture attention.
Approach: They propose a deductive and inductive approach to categorize research impact categories using a corpus-based approach . they use a combination of deductive methods and machine learning to infer impact categories from project reports.
Outcome: The proposed method predicts deductively and inductively derived impact categories with 76.39% accuracy and 78.81% accuracy.
In-depth Research Impact Summarization through Fine-Grained Temporal Citation Analysis (2026.acl-long)

Copied to clipboard

Challenge: citation counts are a shallow view that fails to capture how a paper has influenced subsequent work.
Approach: They propose a task to generate nuanced, expressive, and time-aware impact summaries . they analyze fine-grained confirmatory and correction citation intents to generate summary .
Outcome: The proposed task shows moderate to strong human correlation on subjective metrics such as insightfulness.
CitationIE: Leveraging the Citation Graph for Scientific Information Extraction (2021.acl-long)

Copied to clipboard

Challenge: Existing work on scientific information extraction (SciIE) considers extraction solely based on the content of an individual paper, without considering the paper’s place in the broader literature.
Approach: They propose to automate the extraction of key information from scientific documents by leveraging a complementary source: the citation graph of referential links between citing and cited papers.
Outcome: The proposed model improves on a set of English-language scientific documents.
The Noisy Path from Source to Citation: Measuring How Scholars Engage with Past Research (2025.acl-long)

Copied to clipboard

Challenge: Academic citations are widely used for evaluating research and tracing knowledge flows.
Approach: They propose a computational pipeline to quantify citation fidelity at the sentence level by identifying citations in citing papers and corresponding claims in cited papers.
Outcome: The proposed pipeline identifies citations in citing papers and the corresponding claims in cited papers and applies supervised models to measure fidelity at the sentence level.
SciImpact: A Multi-Dimensional, Multi-Field Benchmark for Scientific Impact Prediction (2026.findings-acl)

Copied to clipboard

Challenge: Prior work on scientific impact prediction has focused on citation counts and its variants, leaving limited evaluation of models’ capability to reason about other dimensions.
Approach: They propose a large-scale, multi-dimensional benchmark for scientific impact prediction spanning 19 fields.
Outcome: The proposed model outperforms larger models and close-source models in a wide range of fields and measures of scientific impact across 19 fields.
Geographic Citation Gaps in NLP Research (2022.emnlp-main)

Copied to clipboard

Challenge: a vast number of papers accepted at top NLP venues come from a handful of western countries and (lately) China.
Approach: They ask researchers to examine the relationship between geographical location and publication success . they use a dataset of 70,000 papers from the ACL Anthology to examine their citation network .
Outcome: The proposed dataset of 70,000 papers from the ACL Anthology shows that there are substantial geographical disparities in paper acceptance and citations .
Examining Citations of Natural Language Processing Literature (2020.acl-main)

Copied to clipboard

Challenge: citations of NLP papers have decreased in recent years, but long papers get three times as many citation as short papers . citation data from the ACL Anthology and Google Scholar can be used to understand the field and quantify the impact of different types of papers.
Approach: They extract data from the ACL Anthology and Google Scholar to examine trends in citations of NLP papers.
Outcome: The results show that only about 56% of the papers in AA are cited ten or more times . CL Journal has the most cited papers, but its citation dominance has lessened .
Beyond Metadata: What Paper Authors Say About Corpora They Use (2021.findings-acl)

Copied to clipboard

Challenge: Currently, dataset retrieval relies almost exclusively on metadata provided by the publishers.
Approach: They propose to use metadata to extract review statements from scientific publications . they argue that a crucial piece of information is missing to inform the examination of search results .
Outcome: The proposed analysis is the first of its kind in the field of Natural Language Processing.
Enhancing Scientific Document Summarization with Research Community Perspective and Background Knowledge (2024.lrec-main)

Copied to clipboard

Challenge: Scientific paper summarization is the focus of recent research . prevailing summarizing methods involve selective extraction of content from abstract, introduction, and conclusion segments within the target articles.
Approach: They propose a model that incorporates references and citations to capture the impact of the document on the research community.
Outcome: The proposed model generates extractive and abstractive summaries in parallel and improves their performance when considering the standard metrics.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations