Automatic identification of writers’ intentions: Comparing different methods for predicting relationship goals in online dating profile texts (D19-55)
Copied to clipboard
| Challenge: | lexicon-based text analysis methods such as LIWC have been criticized by computational linguists for their lack of adaptability, but they have not been systematically compared with either human evaluations or machine learning approaches. |
| Approach: | They used a corpus of online dating profile texts to compare LIWC, machine learning, and a human baseline to assess their effectiveness on a relationship goal classification task. |
| Outcome: | The proposed methods were compared with a corpus of online dating profile texts and a human baseline. |
Similar Papers
A Survey of Automatic Personality Detection from Texts (2020.coling-main)
Copied to clipboard
| Challenge: | Personality profiling has long been used in psychology to predict life outcomes. |
| Approach: | They present the trajectory of automatic personality detection from purely psychology approaches to the latest purely natural language processing approaches on large social media datasets. |
| Outcome: | The proposed models have been compared with the most recent approaches on large social media datasets. |
Cognitive Linguistic Identity Fusion Score (CLIFS): A Scalable Cognition‐Informed Approach to Quantifying Identity Fusion from Text (2025.emnlp-main)
Copied to clipboard
| Challenge: | Existing methods for measuring identity fusion are limited and require controlled surveys or direct field contact. |
| Approach: | They propose a new metric that integrates cognitive linguistics with large language models to measure identity fusion. |
| Outcome: | The proposed metric outperforms existing methods and human annotations in violence risk assessment. |
Exploring Ordinality in Text Classification: A Comparative Study of Explicit and Implicit Techniques (2024.findings-acl)
Copied to clipboard
Siva Rajesh Kasa, Aniket Goel, Karan Gupta, Sumegh Roychowdhury, Pattisapu Priyatam, Anish Bhanushali, Prasanna Srinivasa Murthy
| Challenge: | Ordinal classification (OC) is a key task in natural language processing with applications in various domains such as sentiment analysis, rating prediction, and more. |
| Approach: | They propose to tackle ordinal classification (OC) through the implicit semantics of the labels . they propose to use a classical explicit approach and an implicit approach that organically engages the semantics. |
| Outcome: | The proposed methods are based on pre-trained language models and offer strategic recommendations based upon specific settings. |
The Engage Corpus: A Social Media Dataset for Text-Based Recommender Systems (2022.lrec-1)
Copied to clipboard
| Challenge: | Existing studies have examined the impact of recommendation algorithms on how users discover and join online groups, but there are few standardized datasets for generating such models. |
| Approach: | They propose to use Reddit to build a dataset that can be used to build models of user engagement with online groups. |
| Outcome: | The proposed model is based on the behavior of subreddits banned in June 2020 as part of Reddit's efforts to stop the dissemination of hate speech. |
Beyond Text: Leveraging Multi-Task Learning and Cognitive Appraisal Theory for Post-Purchase Intention Analysis (2024.findings-acl)
Copied to clipboard
| Challenge: | Recent studies have shown that user-level features can carry more task-related information than the text itself. |
| Approach: | They evaluate multi-task learning frameworks grounded in Cognitive Appraisal Theory to predict user behavior as a function of users’ self-expression and psychological attributes. |
| Outcome: | The proposed models improve on the language and traits of users, while lacking rich annotations of other attributes. |
User-Level Race and Ethnicity Predictors from Twitter Text (C18-1)
Copied to clipboard
| Challenge: | Using social media text to identify user-level race and ethnicity is a useful tool for a range of downstream applications, including passive polling or quantifying demographic bias. |
| Approach: | They propose to collect data from social media users who self-report their race/ethnicity through a survey to develop models which accurately predict the membership of a user to the four largest racial and ethnic groups with up to .884 AUC. |
| Outcome: | The proposed models accurately predict the membership of a user to the four largest racial and ethnic groups with up to .884 AUC and make available to the research community. |
Measuring Forecasting Skill from Text (2020.acl-main)
Copied to clipboard
| Challenge: | Prior studies have shown that some individuals can make accurate predictions with consistently better accuracy. |
| Approach: | They examine linguistic factors associated with people's predictions including uncertainty, readability, and emotion. |
| Outcome: | The proposed model can accurately predict forecasting skill using only language. |
Predicting Responses to Psychological Questionnaires from Participants’ Social Media Posts and Question Text Embeddings (2020.findings-emnlp)
Copied to clipboard
| Challenge: | Existing data cannot be used to predict responses for new questions or participants. |
| Approach: | They propose a method that uses social media texts and the text of the question to predict a participant's questionnaire response. |
| Outcome: | The proposed method can be used to integrate new participants or new questions into psychological studies without costly data collection. |
Comparing Text Representations: A Theory-Driven Approach (2021.emnlp-main)
Copied to clipboard
| Challenge: | Recent advances in NLP have been made by learning representations that transform complex tasks into simple classification tasks. |
| Approach: | They propose a method to evaluate the compatibility between representations and tasks by fitting text features to specific characteristics of text datasets. |
| Outcome: | The proposed model provides a calibrated, quantitative measure of the difficulty of a classification-based NLP task. |
‘Am I the Bad One’? Predicting the Moral Judgement of the Crowd Using Pre–trained Language Models (2022.lrec-1)
Copied to clipboard
| Challenge: | Existing studies on NLP touch upon moral contexts in text. |
| Approach: | They construct a dataset that can be used for moral judgement tasks on a popular reddit subreddit. |
| Outcome: | The proposed model passes moral judgements on posts from a popular reddit subreddit . it shows that the model can be fine tuned and improves across the datasets . |