A Survey on Automatically-Constructed WordNets and their Evaluation: Lexical and Word Embedding-based Approaches (L18-1)
Copied to clipboard
| Challenge: | WordNets are lexical databases in which groups of synonyms are stored according to the semantic relationships between them. |
| Approach: | This paper describes various approaches to constructing WordNets automatically by leveraging traditional lexical resources and newer trends such as word embeddings. |
| Outcome: | The proposed methods leverage traditional lexical resources and newer trends such as word embeddings to build and evaluate WordNets. |
Similar Papers
Browsing and Supporting Pluricentric Global Wordnet, or just your Wordnet of Interest (L18-1)
Copied to clipboard
| Challenge: | a wordnet browser that allows to consult wordnet content is presented in this paper . the paper presents a browser that meets design requirements and complies with the most ample range of design features. |
| Approach: | They propose a wordnet browser that meets design requirements for wordnets . they use existing browsers to analyze their functionalities and build a new browser . |
| Outcome: | The proposed browser meets design requirements and complies with the most ample range of design features. |
Introducing Lexical Masks: a New Representation of Lexical Entries for Better Evaluation and Exchange of Lexicons (2020.lrec-1)
Copied to clipboard
| Challenge: | Existing standards for lexicon format and features are inadequate for evaluation and exchange . lexical masks are a powerful tool used to evaluate and exchange large lexiconic databases . |
| Approach: | They propose a tool to evaluate and exchange lexicon databases in many languages . they propose lexical masks which represent the expected internal structure of a lexico . |
| Outcome: | The proposed lexical masks can be used to evaluate and exchange lexicon databases in many languages. |
A Short Survey on Sense-Annotated Corpora (2020.lrec-1)
Copied to clipboard
| Challenge: | Word Sense Disambiguation (WSD) is a key task in Natural Language Understanding. |
| Approach: | They propose to use sense-annotated corpora for supervised Word Sense Disambiguation. |
| Outcome: | The proposed methods have been compared with knowledge-based approaches and have shown to be more efficient when they are available. |
Latent semantic network induction in the context of linked example senses (D19-55)
Copied to clipboard
| Challenge: | Using the Princeton WordNet, we construct a network using the entirety of Wiktionary. |
| Approach: | They propose to use Wiktionary to construct a wordnet using the entirety of the open-source dictionary. |
| Outcome: | The proposed network induction process is similar to the Princeton WordNet, but with a more data-driven approach. |
The interplay between lexical resources and Natural Language Processing (N18-6)
Copied to clipboard
| Challenge: | linguistic, world and common sense knowledge is an important research area, but processing and storing it in lexical resources is not a straightforward task. |
| Approach: | They propose to use NLP methods to help process of constructing and enriching lexical resources and the use of lexicals for improving NLP applications. |
| Outcome: | The proposed approach aims to speed up and/or ease up the process of resource curation and enrichment. |
Some Issues with Building a Multilingual Wordnet (2020.lrec-1)
Copied to clipboard
| Challenge: | Notable extensions include: confidence, corpus frequency, orthographic variants, lexicalized and non-lexicalised synsets and lemmas, new parts of speech, and more. |
| Approach: | They propose to integrate a new open multilingual wordnet format that tests the extensions introduced by the new format and integrates a set of tools to ensure the integrity of the Collaborative Interlingual Index. |
| Outcome: | The proposed format integrates a set of tools that test the extensions while ensuring the integrity of the Collaborative Interlingual Index (CILI). |
Using Wiktionary to Create Specialized Lexical Resources and Datasets (2022.lrec-1)
Copied to clipboard
| Challenge: | Using Wiktionary data to build specialized lexical datasets can be used for evaluating or improving NLP tasks, like Word Sense Disambiguation (WSD), Word-in-Context challenges (WiC), or Machine Translation (MT). |
| Approach: | They propose to use Wiktionary data to create specialized lexical datasets that can be used for evaluating or improving NLP tasks. |
| Outcome: | The proposed datasets can be used to improve and/or evaluate NLP tasks, like Word Sense Disambiguation (WSD), Word-in-Context challenges (WiC), or Sense Linking (SL), or machine translation (MT). |
A Method for Studying Semantic Construal in Grammatical Constructions with Interpretable Contextual Embedding Spaces (2023.acl-long)
Copied to clipboard
| Challenge: | Existing paradigms for the linguistically oriented exploration of large neural language models include treating the model as a linguistic test subject by measuring output on test sentences and building probing classifiers on top of embeddings to test whether the embeddables are sensitive to certain properties like dependency structure. |
| Approach: | They project contextual embeddings into interpretable semantic spaces, each defined by a different set of psycholinguistic feature norms. |
| Outcome: | The proposed method can probe the distributional meaning of syntactic constructions at a templatic level, abstracted away from specific lexemes. |
Analyzing the Surprising Variability in Word Embedding Stability Across Languages (2021.emnlp-main)
Copied to clipboard
| Challenge: | Word embeddings are powerful representations that form the foundation of many natural language processing architectures. |
| Approach: | They explore word embedding stability in a wide range of languages to gain insight into their stability. |
| Outcome: | The proposed results provide insights into word embedding stability in English and other languages. |
Inferences for Lexical Semantic Resource Building with Less Supervision (2020.lrec-1)
Copied to clipboard
| Challenge: | lexical semantic resources may be built using various approaches such as extraction from corpora, integration of relevant pieces of knowledge from pre-existing knowledge resources and endogenous inference. |
| Approach: | They propose a method where the resource building process appears as a self learning process . they propose lexical and semantic resource building based on inference . |
| Outcome: | The proposed method reduces the human effort needed for lexical semantic resource building. |