Challenge: Terminology standardization plays an important role in the management of terminological resources.
Approach: They propose to re-model an existing multilingual terminological database for the medical domain, TriMED, and propose a method to make it compliant to the latest ISO/TC 37 standards.
Outcome: The proposed model should be compliant with the three most recent ISO/TC 37 standards and has a new data category repository and a Web application that can be used to access the multilingual terminological records.

Similar Papers

TriMED: A Multilingual Terminological Database (L18-1)

Copied to clipboard

Challenge: a terminological tool is developed to solve communication problems in medical language . medical terminology is often semantically complex and difficult to understand .
Approach: They propose a terminological tool that solves problems related to the opacity of medical language . they use a multilingual terminological-phraseological resource called TriMED to analyze medical terminology .
Outcome: The proposed tool solves problems related to the opacity that characterizes communication in the medical field among its actors.
From Linguistic Resources to Ontology-Aware Terminologies: Minding the Representation Gap (2020.lrec-1)

Copied to clipboard

Challenge: Terminological resources are not available in standard formats such as Term Base eXchange (TBX) thus preventing their sharing and reuse.
Approach: They propose to convert terminological resources into TBX format and to integrate ontology-based information into terminologies.
Outcome: The proposed tool supports the process of creating ontology-aware terminologies . terminologie creation and maintenance determine the quality of the final product of a translation process .
Creating Terminological Resources in the Digital Age for Less-resourced Languages (2024.lrec-main)

Copied to clipboard

Challenge: Multilingual terminological resources are limited in less resourced languages, limiting knowledge spread in less-resourced languages . linguists and terminologists must use natural language processing tools to maximize resources . less-represented languages suffer from a lack of available linguistic resources - a survey shows .
Approach: They propose a method to maximize the open access catalan terminology available . authors propose linguists and terminologists supervise the project and translate it into catalane .
Outcome: The proposed method maximizes the catalan terminology currently available in open access . the results are supervised by linguists and terminologists experts before being publicly available to the public.
Multilingualization of Medical Terminology: Semantic and Structural Embedding Approaches (2020.lrec-1)

Copied to clipboard

Challenge: Existing methods for multilingual terminology curation are limited as they do not fit the term within existing terminology.
Approach: They propose a method to encode the structural property of a term by aligning embeddings using graph convolutional networks trained from separate languages.
Outcome: The proposed method can encode the structural property of a term by aligning embeddings using graph convolutional networks trained from separate languages.
An Integrated Formal Representation for Terminological and Lexical Data included in Classification Schemes (L18-1)

Copied to clipboard

Challenge: e-lexicography is a field of study dealing with the automated creation of specialized multilingual dictionaries from structured data.
Approach: They propose to use a SKOS-XL vocabulary for modelling the multilingual terminological part of comparable taxonomies and OntoLex-Lemon for modelling multilingual lexical entries.
Outcome: The proposed model can be explicitly cross-linked in the context of the Linguistic Linked Open Data (LLOD).
A Gold Standard for Multilingual Automatic Term Extraction from Comparable Corpora: Term Structure and Translation Equivalents (L18-1)

Copied to clipboard

Challenge: Terms are notoriously difficult to identify, both automatically and manually.
Approach: They propose a method to annotate terms manually from a comparable corpus . they show that the gold standard provides a tool for evaluation and a rich source of information .
Outcome: The proposed method provides a tool for evaluation and rich source of information about terms.
Representing Multiword Term Variation in a Terminological Knowledge Base: a Corpus-Based Study (2020.lrec-1)

Copied to clipboard

Challenge: Multiword terms are the most frequent type of lexical units in scientific and technical communication. rendering them in another language is not easy due to their cognitive complexity, proliferation of different forms, and their unsystematic representation in terminographic resources.
Approach: They evaluated Spanish translation variants of multiword terms in three parallel corpora, two comparable corporales and two terminological resources.
Outcome: The results show that multiword terms exhibit a significant degree of term variation . the proposed model is based on a set of criteria for determining which variants should be selected .
Leveraging Domain Corpora for Enhanced Terminology: The Case of Estonian-English Remote Sensing Termbase (2024.lrec-main)

Copied to clipboard

Challenge: Termbase is a domain corpora and terminological database for remote sensing in Estonia.
Approach: They propose to develop an Estonian-English Remote Sensing Termbase from scratch . they use the Estonian Remote Sensenting Corpus 2022 as the primary data source .
Outcome: The Estonian Remote Sensing Corpus 2022 served as the primary data source for the termbase.
A Taxonomy for In-depth Evaluation of Normalization for User Generated Content (L18-1)

Copied to clipboard

Challenge: Existing taxonomies for lexical normalization are not suitable for the task of normalization since the categories are substantially different.
Approach: They propose a taxonomy of error categories for lexical normalization . they annotate a recent normalization dataset and read a near-perfect agreement .
Outcome: The proposed taxonomy is based on a recent normalization dataset and it performs well.
An LLM-based Framework for Biomedical Terminology Normalization in Social Media via Multi-Agent Collaboration (2025.coling-main)

Copied to clipboard

Challenge: Experimental results indicate that our approach exhibits competitive performance.
Approach: They propose a tuning-free approach to normalize non-standard terms using large language models . they use a search engine and a domain knowledge base to expand the short texts into accurate descriptions .
Outcome: The proposed approach is based on the "Recall and Re-rank" framework . it can be used to identify the standard term in a specified termbase for non-standardized mentions .

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations