Challenge: Existing methods for interactive dictionary construction are limited to a small number of terms, but we propose a method that can be used to create flexible dictionaries with precise granularity.
Approach: They propose a method to construct a term dictionary for text analytics through an interactive process between a human and a machine.
Outcome: The proposed method outperforms baseline methods and works even with a small number of interactions.

Similar Papers

lingvis.io - A Linguistic Visual Analytics Framework (P19-3)

Copied to clipboard

Challenge: Using a modular framework, linguistic visual analytics applications can be rapidly prototypized using a web-based framework.
Approach: They propose a modular framework for rapid prototyping of linguistic, web-based, visual analytics applications.
Outcome: The proposed framework supports rapid prototyping of linguistic, web-based, visual analytics applications.
Lexi: A tool for adaptive, personalized text simplification (C18-1)

Copied to clipboard

Challenge: Existing research on text simplification has aimed to develop generic solutions . instead, we need to develop customized simplification systems for individual users .
Approach: They propose a framework for adaptive lexical simplification and introduce Lexi, a free open-source tool for personalized text simplification.
Outcome: The proposed framework is based on a free open-source tool for adaptive, personalized text simplification.
A Survey on Automatically-Constructed WordNets and their Evaluation: Lexical and Word Embedding-based Approaches (L18-1)

Copied to clipboard

Challenge: WordNets are lexical databases in which groups of synonyms are stored according to the semantic relationships between them.
Approach: This paper describes various approaches to constructing WordNets automatically by leveraging traditional lexical resources and newer trends such as word embeddings.
Outcome: The proposed methods leverage traditional lexical resources and newer trends such as word embeddings to build and evaluate WordNets.
An LLM-Based Approach for Insight Generation in Data Analysis (2025.naacl-long)

Copied to clipboard

Challenge: Existing approaches to generate insightful data from databases are time-consuming and resource-intensive.
Approach: They propose a method that leverages Large Language Models to automatically generate textual insights from databases.
Outcome: The proposed approach generates more insightful insights than other approaches while maintaining correctness.
Text Mining for History: first steps on building a large dataset (L18-1)

Copied to clipboard

Challenge: a new corpus on the history domain is being created to mine text in the domain . primary motivation for the project is the need to query the material in a non-linear way .
Approach: They propose to use a Brazilian historical-biographical dictionary as a resource for text mining.
Outcome: The proposed corpus is a reference work on the Brazilian history domain . it contains almost 12 millions tokens in about three hundred thousand sentences . the authors argue that the proposed corpu is linguistically motivated .
ScheMatiQ: From Research Question to Structured Data through Interactive Schema Discovery (2026.acl-demo)

Copied to clipboard

Challenge: a new approach to natural-language research questions requires manual effort to generate an annotation schema and label the corpus.
Approach: They propose a natural-language search tool that takes a question and a corpus to produce a schema and db with a web interface that lets steer and revise the extraction.
Outcome: The proposed model yields outputs that support real-world analysis in law and computational biology.
TinyScientist: An Interactive, Extensible, and Controllable Framework for Building Research Agents (2025.emnlp-demos)

Copied to clipboard

Challenge: Existing research systems often design and use agentic workflows to perform research tasks such as ideation, scientific coding, review writing, and tree-based search.
Approach: They propose an open-source codebase, an interactive web demonstration, and a PyPI Python package to make state-of-the-art auto-research pipelines broadly accessible to every researcher and developer.
Outcome: The proposed framework adapts easily to new tools and supports iterative growth.
Event-Centric Natural Language Processing (2021.acl-tutorials)

Copied to clipboard

Challenge: This tutorial will provide an introduction to various methods for automating the extraction, conceptualization and prediction of events and their relations.
Approach: This tutorial will provide an introduction to various methods for automating events and their relations, and a wide range of NLU and commonsense understanding tasks.
Outcome: This tutorial will provide an introduction to various methods for automating extraction, conceptualization and prediction of events and their relations, and a wide range of NLU and commonsense understanding tasks.
LUCE: A Dynamic Framework and Interactive Dashboard for Opinionated Text Analysis (2025.coling-demos)

Copied to clipboard

Challenge: LUCE is an advanced dynamic framework for analysing opinionated text . it features computational modules for different elements of opinions, e.g., sentiment/emotion, suggestion, figurative language, hate/toxic speech, and topics.
Approach: They introduce a dynamic framework with an interactive dashboard for analysing opinionated text . it features computational modules of text classification and extraction for different elements of opinions .
Outcome: The framework is validated in a relevant environment and its capabilities and performance demonstrated . it features trained models, python-based APIs, and a user-friendly dashboard .
AnnoPlot: Interactive Visualizations of Text Annotations (2024.eacl-demo)

Copied to clipboard

Challenge: Annotation projects face challenges in data quality and validity, authors argue .
Approach: They propose an open-source web application that analyzes, manages, and visualizes annotated text data.
Outcome: The proposed application is open-source and promotes transparency and user control . it offers comprehensive views of span annotations and category systems without training or classification model .

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations