| Challenge: | a scholarly retweeting prediction system is proposed to predict scholarly tweets . re-tweening is an action of reposting others' tweet by using the reretwet button on Twitter . |
| Approach: | They propose a real-time scholarly retweeting prediction system that retrieves scholarly tweets which will be re-tweeled. |
| Outcome: | The proposed system outperforms baseline systems and can predict scientific impact in real-time. |
Similar Papers
A Real-Time System for Credibility on Twitter (2020.lrec-1)
Copied to clipboard
| Challenge: | Using neural networks, we can analyze Twitter in real-time to determine whether users are credible and false. |
| Approach: | They propose to analyze Twitter in real-time using neural networks to determine credibility of tweets and users who posted them. |
| Outcome: | The proposed method analyzes Twitter in real-time to determine which users are credible and which are not, what is false or what is true on the Internet. |
Realistic Citation Count Prediction Task for Newly Published Papers (2023.findings-eacl)
Copied to clipboard
| Challenge: | Existing studies on citation count prediction assume that future citation counts of academic papers have not had enough time pass since publication. |
| Approach: | They propose to use citation counts of newly published papers as a realistic citation count prediction task and to use them to leverage the citations of papers shortly after publication. |
| Outcome: | The proposed methods significantly improve the performance of citation count prediction for newly published papers in a realistic setting. |
Automatic Generation of Citation Texts in Scholarly Papers: A Pilot Study (2020.acl-main)
Copied to clipboard
| Challenge: | Existing studies on automatic generation of citation texts in scholarly papers have not investigated this problem. |
| Approach: | They propose to train an implicit citation extraction model based on BERT and a multi-source pointer-generator network with cross attention mechanism for citation text generation. |
| Outcome: | The proposed model can generate short texts to describe cited papers in scholarly papers with training data. |
Predicting Factuality of Reporting and Bias of News Media Sources (D18-1)
Copied to clipboard
| Challenge: | a new study examines the factuality of news media and its biases . social media has democratized content creation and spread information online . |
| Approach: | They propose to characterize entire news media to predict factuality and bias . they experiment with news websites and a set of features derived from their content . |
| Outcome: | The proposed model shows that the features of news websites perform better than baseline . the results show that the feature types are important for fact-checking systems . |
SEDTWik: Segmentation-based Event Detection from Tweets Using Wikipedia (N19-3)
Copied to clipboard
| Challenge: | Recent work on event detection from tweets has focused on localized events or breaking news only. |
| Approach: | They propose to split tweets into segments, extract bursty segments, cluster them, summarize them. |
| Outcome: | The proposed system can detect newsworthy events occurring at different locations of the world from a wide range of categories. |
Prediction for the Newsroom: Which Articles Will Get the Most Comments? (N18-3)
Copied to clipboard
| Challenge: | a new method to support manual moderation of discussion sections is proposed. |
| Approach: | They propose to support manual moderation by proactively drawing attention of moderators to articles that most likely need their intervention. |
| Outcome: | The proposed method outperforms the current state-of-the-art methods on a 7-million-comment dataset. |
‘Quis custodiet ipsos custodes?’ Who will watch the watchmen? On Detecting AI-generated peer-reviews (2024.emnlp-main)
Copied to clipboard
| Challenge: | Recent studies have focused on generic AI-generated text detection or estimating fraction of peer-reviews that can be AI-generated. |
| Approach: | They propose a model that detects whether a peer-review is written by ChatGPT and a reviewer-generated model that generates similar outputs upon re-prompting. |
| Outcome: | The proposed model is more robust, but paraphrasing is more effective. |
The Engage Corpus: A Social Media Dataset for Text-Based Recommender Systems (2022.lrec-1)
Copied to clipboard
| Challenge: | Existing studies have examined the impact of recommendation algorithms on how users discover and join online groups, but there are few standardized datasets for generating such models. |
| Approach: | They propose to use Reddit to build a dataset that can be used to build models of user engagement with online groups. |
| Outcome: | The proposed model is based on the behavior of subreddits banned in June 2020 as part of Reddit's efforts to stop the dissemination of hate speech. |
TweetNLP: Cutting-Edge Natural Language Processing for Social Media (2022.emnlp-demos)
Copied to clipboard
Jose Camacho-collados, Kiamehr Rezaee, Talayeh Riahi, Asahi Ushio, Daniel Loureiro, Dimosthenis Antypas, Joanne Boisson, Luis Espinosa Anke, Fangyu Liu, Eugenio Martínez Cámara
| Challenge: | TweetNLP is an integrated platform for natural language processing in social media. |
| Approach: | They propose a Python-based platform for natural language processing in social media that supports a variety of NLP tasks. |
| Outcome: | The proposed platform supports generic focus areas such as sentiment analysis and named entity recognition, as well as social media-specific tasks such as emoji prediction and offensive language identification. |
Rumor Detection on Social Media: Datasets, Methods and Opportunities (D19-50)
Copied to clipboard
| Challenge: | Social media platforms are used for information gathering, but they also lead to the spreading of rumors and fake news. |
| Approach: | This paper presents a comprehensive list of datasets used for rumor detection . it also reviews the important studies based on what types of information they exploit . |
| Outcome: | This paper presents an overview of the recent studies in the rumor detection field . it provides a comprehensive list of datasets used for rumour detection . |