| Challenge: | Existing lexical resources for semantic annotation of synonyms are lacking in computational language processing. |
| Approach: | They describe a bilingual lexical resource being built to investigate verbal synonymy in bilingual context and relate semantic roles common to one synonym class to verb arguments. |
| Outcome: | The proposed resource is based on English and Czech WordNet, FrameNet, PropBank, VerbNet (SemLink), and valency lexicons for Czech and English (PDT-Vallex, Vallex, and EngValleX). |
Similar Papers
Creating a Verb Synonym Lexicon Based on a Parallel Corpus (L18-1)
Copied to clipboard
| Challenge: | a new lexical resource called CzEngClass is being built to help define synonyms in a bilingual context. |
| Approach: | They propose to group verb senses into bilingual verbal synonym groups and use a parallel dependency corpus to explore semantic 'equivalence' they argue that existence of core argument mappings and adjunct mappings to a common set of semantic roles is a suitable criterion for a reasonable verb synonymy definition . |
| Outcome: | The proposed resource will be available by mid-2018 . |
Tools for Building an Interlinked Synonym Lexicon Network (L18-1)
Copied to clipboard
| Challenge: | a new lexicon is being developed for cross-lingual (Czech and English) synonyms based on their syntactic and semantic behavior in (bilingual) context. |
| Approach: | They propose to build a new interlinked verbal synonym lexicon called CzEngClass using a tool that helps to keep cross-lingual synonym classes consistent. |
| Outcome: | The proposed lexicon captures cross-lingual (Czech and English) synonyms . the tool, called Synonym Class Editor -SynEd, is customized to build and edit entries . |
GeCzLex: Lexicon of Czech and German Anaphoric Connectives (2020.lrec-1)
Copied to clipboard
| Challenge: | Existing lexicons of connectives are interlinked with each other to provide a bilingual inventory of connective entries. |
| Approach: | They introduce the first version of a lexicon for translation equivalents of Czech and German discourse connectives. |
| Outcome: | The lexicon is the first bilingual inventory of connectives with linkage on the level of individual entries. |
NomVallex: A Valency Lexicon of Czech Nouns and Adjectives (2022.lrec-1)
Copied to clipboard
| Challenge: | NomVallex is a manual annotated valency lexicon of Czech nouns and adjectives . valencies are the ability of a verb to combine with other sentence constituents based on their morphemic forms . |
| Approach: | They propose a manually annotated valency lexicon of Czech nouns and adjectives . they capture valencies of a lexical unit in a sequence of valence slots . |
| Outcome: | The proposed lexicon is based on corpus data and contains 1027 lexical units . valency properties of lexicals are captured in a valence frame, with morphemic forms . |
A Survey on Automatically-Constructed WordNets and their Evaluation: Lexical and Word Embedding-based Approaches (L18-1)
Copied to clipboard
| Challenge: | WordNets are lexical databases in which groups of synonyms are stored according to the semantic relationships between them. |
| Approach: | This paper describes various approaches to constructing WordNets automatically by leveraging traditional lexical resources and newer trends such as word embeddings. |
| Outcome: | The proposed methods leverage traditional lexical resources and newer trends such as word embeddings to build and evaluate WordNets. |
Exploring the Representation of Word Meanings in Context: A Case Study on Homonymy and Synonymy (2021.acl-long)
Copied to clipboard
| Challenge: | Existing models that represent different senses of words in context are not accurate for polysemous words. |
| Approach: | They propose a multilingual dataset that evaluates the ability of models to accurately represent different lexical-semantic relations such as homonymy and synonymy. |
| Outcome: | The proposed models can disambiguate homonyms in context, but fail to represent words with different senses when occurring in similar sentences. |
Making a Semantic Event-type Ontology Multilingual (2022.lrec-1)
Copied to clipboard
| Challenge: | a new version of SynSemClass is being developed for use in natural language processing . the ontology is a bilingual resource with no links to a valency lexicon . |
| Approach: | They propose to add German entries to the SynSemClass Event-type Ontology . they propose to use the ontology as a human-readable and human-understandable database . |
| Outcome: | The proposed extension of SynSemClass Event-type Ontology is presented in a paper in czech republic . the ontology provides curated data for NLP experiments with cross-lingual synonyms . |
Annotating the French Wiktionary with supersenses for large scale lexical analysis: a use case to assess form-meaning relationships within the nominal lexicon (2025.coling-main)
Copied to clipboard
| Challenge: | Conducting large-scale empirical studies in lexical semantics remains an elusive goal for many languages lacking comprehensive semantic resources. |
| Approach: | They propose to use the Princeton WordNet to enrich the French Wiktionary with general semantic classes, known as supersenses, using a limited amount of manually annotated data. |
| Outcome: | The proposed method can be extended to other languages provided an electronic lexicon and manually annotated senses are available. |
A Dataset of Translational Equivalents Built on the Basis of plWordNet-Princeton WordNet Synset Mapping (2020.lrec-1)
Copied to clipboard
| Challenge: | a dataset of 11,000 Polish-English translational equivalents is presented . the dataset is a novum in the wordnet domain and can facilitate the precision of bilingual NLP tasks. |
| Approach: | They present a dataset of Polish-English translational equivalents linked by three types of equivalence links. |
| Outcome: | The proposed dataset contains 11,000 Polish-English translational equivalents . the resulting subsets are based on a manual annotation process and a set of formal features . |
Prague Dependency Treebank - Consolidated 1.0 (2020.lrec-1)
Copied to clipboard
Jan Hajič, Eduard Bejček, Jaroslava Hlavacova, Marie Mikulová, Milan Straka, Jan Štěpánek, Barbora Štěpánková
| Challenge: | Using the standard PDT scheme, the Prague Dependency Treebank-Consolidated 1.0 contains 4 different datasets of Czech, uniformly annotated using the standard scheme. |
| Approach: | They present a richly annotated and genre-diversified language resource, the Prague Dependency Treebank-Consolidated 1.0, which contains 4 different datasets of Czech, uniformly annnotated using the standard PDT scheme. |
| Outcome: | The Prague Dependency Treebank-Consolidated 1.0 contains 4 datasets of Czech, uniformly annotated using the standard PDT scheme. |