Orchestrating NLP Services for the Legal Domain (2020.lrec-1)

Copied to clipboard

Challenge: a legal technology system under development in the EU is based on semantic services and a multilingual legal knowledge Graph.
Approach: They propose a workflow manager that enables flexible orchestration of workflows . they describe different use cases and propose prototypical solutions .
Outcome: The proposed system is based on a set of natural language processing and document curation services and a multilingual legal knowledge graph that contains semantic information and meaningful references to legal documents.

Similar Papers

LAW: Legal Agentic Workflows for Custody and Fund Services Contracts (2025.coling-industry)

Copied to clipboard

Challenge: Currently, there are limited resources available to build a legal domain-specific Large Language Model (LLM) however, legal contracts are highly varied not only in terms of semantics but also accessibility.
Approach: They propose a Large Language Model (LLM) that integrates multiple specialized agents and text agents to respond to user queries.
Outcome: The proposed model outperforms the baseline model in complex tasks such as calculating a contract’s termination date by 92.9% points.
The Law and NLP: Bridging Disciplinary Disconnects (2023.findings-emnlp)

Copied to clipboard

Challenge: Legal practitioners and scholars have been slow to adopt tools from natural language processing (NLP) the legal system is experiencing an access to justice crisis, which could be partially alleviated with NLP.
Approach: They argue that legal practitioners are slow to adopt natural language processing (NLP) they argue that there is a disconnect between legal needs and NLP research .
Outcome: The proposed tasks bridge disciplinary disconnects and highlight interesting areas for legal NLP research that remain underexplored.
Towards Automated Extraction of Business Constraints from Unstructured Regulatory Text (C18-2)

Copied to clipboard

Challenge: a system for machine-driven annotations of legal documents is currently undergoing user trials within our organization.
Approach: a system for machine-driven annotations of legal documents is presented . the system is currently undergoing user trials within our organization.
Outcome: the proposed system is currently undergoing user trials within our organization.
LEXTREME: A Multi-Lingual and Multi-Task Benchmark for the Legal Domain (2023.findings-emnlp)

Copied to clipboard

Challenge: Recent advances in legal NLP have led to a rapid growth of the field . however, many benchmarks are available only in English and no multilingual benchmark exists .
Approach: They propose to use 11 datasets covering 24 languages to compare NLP models.
Outcome: The proposed benchmarks show that even the best baseline only achieves modest results and ChatGPT struggles with many tasks.
LEGAL-BERT: The Muppets straight out of Law School (2020.findings-emnlp)

Copied to clipboard

Challenge: Existing guidelines for pre-training and fine-tuning do not always generalize well in the legal domain.
Approach: They propose to use BERT out of the box, adapt it by additional pre-training on domain-specific corpora, and pre-train it from scratch on domains.
Outcome: The proposed strategies are: use the original BERT out of the box, adapt it by additional pre-training on domain-specific corpora, and pre-train it from scratch on domain specific corpors.
Populating Legal Ontologies using Semantic Role Labeling (2020.lrec-1)

Copied to clipboard

Challenge: This paper is concerned with the ‘resource consumption bottleneck’ of creating semantic technologies manually.
Approach: They propose to combine general-purpose NLP modules with pre- and post-processing using rules based on domain knowledge to solve the acquisition paradox.
Outcome: The proposed system extracts norms from legislation and represents them as structured norms in legal ontologies.
BriefMe: A Legal NLP Benchmark for Assisting with Legal Briefs (2025.findings-acl)

Copied to clipboard

Challenge: a core part of legal work that has been underexplored in Legal NLP is the writing and editing of legal briefs.
Approach: They propose to use large language models to help legal professionals with writing briefs by capturing and evaluating their abilities in language models.
Outcome: The proposed tasks show that the models perform well on arguments summarization, argument completion, and case retrieval tasks.
LexGLUE: A Benchmark Dataset for Legal Language Understanding in English (2022.acl-long)

Copied to clipboard

Challenge: Laws and their interpretations, legal arguments and agreements are typically expressed in writing.
Approach: They propose a benchmark to evaluate model performance across legal NLU tasks . they also evaluate several generic and legal-oriented models .
Outcome: The proposed model performs better across multiple tasks than previous models.
Automated Refugee Case Analysis: A NLP Pipeline for Supporting Legal Practitioners (2023.findings-acl)

Copied to clipboard

Challenge: In Canada, retrieving similar cases and their analysis is a key part of legal work . long processing times are due to a significant backlog and to the amount of work required from counsels .
Approach: They propose to extend existing neural named-entity recognition models to retrieve 19 categories of items from refugee cases.
Outcome: The proposed pipeline achieves a superior F1- score on five of the targeted categories and superior to 80% on an additional 4 categories.
FourCorners: A Production Knowledge Graph Unifying Thailand’s Legal System (2026.acl-industry)

Copied to clipboard

Challenge: Thai legal data lacks standardized, machine-readable data formats . authors: combining legal data requires understanding structural relationships that no existing resource captures.
Approach: They propose a unified temporal knowledge graph for Thai legal data . it integrates 3,840 laws with 87,394 Supreme Court decisions, updated daily .
Outcome: The proposed graph integrates 3,840 laws with 87,394 Supreme Court decisions . it achieves Citation F1 of 0.812 versus 0.666 for practitioner-standard web search .

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations