Multi-Operational Mathematical Derivations in Latent Space (2024.naacl-long)

Copied to clipboard

Challenge: Using a symbolic engine, we investigate the possibility of approximating multiple mathematical operations in latent space for expression derivation.
Approach: They propose to model mathematical operations as explicit geometric transformations by leveraging a symbolic engine and a large-scale dataset.
Outcome: The proposed paradigms can be used to approximate multiple mathematical operations in latent space, while discriminating the conclusions for a single operation is achievable in the original expression encoder.

Similar Papers

MathQA: Towards Interpretable Math Word Problem Solving with Operation-Based Formalisms (N19-1)

Copied to clipboard

Challenge: Existing datasets in this domain do not offer precise operational annotations over diverse problem types due to noise and lack of formal operation-based representations.
Approach: They propose a representation language to map problems to their operation programs . they also introduce an interpretable neural math problem solver .
Outcome: The proposed model outperforms baseline models and the AQUA-RAT dataset on the AQuA-rat dataset.
An Expression Tree Decoding Strategy for Mathematical Equation Generation (2023.emnlp-main)

Copied to clipboard

Challenge: Existing approaches to generate mathematical equations from natural language ignore parallel or dependent relations between math expressions.
Approach: They propose to integrate tree structure into the expression-level generation and advocate an expression tree decoding strategy.
Outcome: The proposed method outperforms baseline methods for generating mathematical equations from natural language.
Rethinking Action Spaces for Reinforcement Learning in End-to-end Dialog Agents with Latent Variable Models (N19-1)

Copied to clipboard

Challenge: Existing approaches to define action spaces for conversational agents have limitations . end-to-end dialog systems can handle complex domains with limited action space .
Approach: They propose a latent action framework that treats the action spaces of an end-to-end dialog agent as latent variables and develops unsupervised methods to induce its own action space from the data.
Outcome: The proposed framework achieves better performance than word-level policy gradient methods on DealOrNoDeal and MultiWoz dialogs.
Representing Syntax and Composition with Geometric Transformations (2021.findings-acl)

Copied to clipboard

Challenge: Existing models of word meaning are based on syntactic rather than proximal co-occurrences, but they are not suitable for syntax sensitive composition.
Approach: They propose to encode syntactic structure by extending the Skip-Gram with Negative sampling architecture from word2vec.
Outcome: The proposed models perform favourably on benchmark word similarity tasks on similarity tests on similar words compared to models based on proximal co-occurrence . however, the real promise of distributional models is the potential for syntax-sensitive composition.
AdaPT: A Set of Guidelines for Hyperbolic Multimodal Multilingual NLP (2024.findings-naacl)

Copied to clipboard

Challenge: Euclidean space is used for training neural models and performing arithmetic operations, but many data types have complex geometries and cannot be captured in the Euclidesan space.
Approach: They propose a set of guidelines for initialization, parametrization, and training of neural networks that can be generalized over existing neural network training methodologies.
Outcome: The proposed framework outperforms Euclidean methods on three tasks over 12 languages and modalities on a variety of domains.
\mathcal XFT: Unlocking the Power of Code Instruction Tuning by Simply Merging Upcycled Mixture-of-Experts (2024.acl-long)

Copied to clipboard

Challenge: Existing studies focus on the data perspectives of instruction tuning, leaving room for exploring advanced training schemes.
Approach: They argue that prior works overlook the possibility of improving code instruction tuning by advancing existing training schemes.
Outcome: The proposed model is dense because all parameters are activated to predict the next token (assuming it is a decoder-only LLM).
Generating Equation by Utilizing Operators : GEO model (2020.coling-main)

Copied to clipboard

Challenge: Existing neural models that use hand-crafted features are expensive and lack domain-specific knowledge.
Approach: They propose a GEO model that uses operator-based features to generate equations using natural language sentences.
Outcome: The proposed model outperforms state-of-the-art models on two datasets and 82.1% in ALG514.
Erratum: Measuring and Improving Consistency in Pretrained Language Models (2021.tacl-1)

Copied to clipboard

Challenge: During production of this paper, an error was introduced to the formula on the bottom of the right column of page 1020.
Approach: the formula was changed in the last two terms of the paper .
Outcome: the correct formula is now available on the web.
Compounding Geometric Operations for Knowledge Graph Completion (2023.acl-long)

Copied to clipboard

Challenge: Knowledge graph embedding (KGE) is one of the most fundamental problems in AI research.
Approach: They propose a new knowledge graph embedding model by leveraging translation, rotation, and scaling operations to form a composite one.
Outcome: The proposed model outperforms existing models on three KG prediction tasks.
Proceedings of the 3rd Workshop on Neural Generation and Translation (D19-56)

Copied to clipboard

Challenge: The third workshop on neural generation and translation is held in london . the workshop received 68 submissions from leading minds in the field .
Approach: the third workshop on neural generation and translation is held in london . the workshop will feature four invited talks from leading minds in the field .
Outcome: the third workshop on neural generation and translation is held in london . the conference received 68 submissions from which 36 accepted .

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations