STAR: Cross-modal [STA]tement [R]epresentation for selecting relevant mathematical premises (2021.eacl-main)
Copied to clipboard
| Challenge: | Existing representations of mathematical statements in natural language are ineffective . STAR model uses cross-modal attention to represent mathematical text . |
| Approach: | They propose a model that uses cross-modal attention to represent mathematical text . it uses conjectures written in both natural and mathematical language to recommend premises . |
| Outcome: | The proposed model outperforms baseline models that do not distinguish between natural and mathematical elements and achieves better performance than state-of-the-art models. |
Similar Papers
Natural Language Premise Selection: Finding Supporting Statements for Mathematical Text (2020.lrec-1)
Copied to clipboard
| Challenge: | Existing approaches to understand mathematical discourse are limited by the complexity of word and symbol interactions. |
| Approach: | They propose a task to retrieve supporting definitions and supporting propositions from a dataset that can be used to evaluate different approaches for the task. |
| Outcome: | The proposed task is based on a dataset that can be used to evaluate different approaches for the natural premise selection task. |
Premise Selection in Natural Language Mathematical Texts (2020.acl-main)
Copied to clipboard
| Challenge: | Existing tasks for natural language premise selection are limited and difficult for humans to interpret and write. |
| Approach: | They propose to use natural language premise selection task to predict premises that will be useful to prove a particular statement. |
| Outcome: | The proposed approach improves the performance of baselines and multi-hop premise selection tasks. |
Introduction to Mathematical Language Processing: Informal Proofs, Word Problems, and Supporting Tasks (2023.tacl-1)
Copied to clipboard
| Challenge: | Using mathematical language processing methods, we analyze prevailing methods, existing limitations, and promising avenues for future research. |
| Approach: | They analyze mathematical language processing methods from recent years and highlight prevailing methodologies, existing limitations and promising avenues for future research. |
| Outcome: | The proposed methods highlight prevailing methods, existing limitations and promising avenues for future research. |
MathAlign: Linking Formula Identifiers to their Contextual Natural Language Descriptions (2020.lrec-1)
Copied to clipboard
Maria Alexeeva, Rebecca Sharp, Marco A. Valenzuela-Escárcega, Jennifer Kadowaki, Adarsh Pyarelal, Clayton Morrison
| Challenge: | Existing approaches to extract mathematical concepts and their descriptions are useful for a variety of tasks, including math information retrieval and accessibility efforts to make scientific documents available to the visually impaired. |
| Approach: | They propose a rule-based approach which extracts LaTeX representations of formula identifiers and links them to their in-text descriptions, given only the original PDF and the location of the formula of interest. |
| Outcome: | The proposed approach extracts LaTeX representations of formula identifiers and links them to their in-text descriptions, given only the original PDF and the location of the formula of interest. |
STEM-POM: Evaluating Language Models Math-Symbol Reasoning in Document Parsing (2025.findings-acl)
Copied to clipboard
| Challenge: | Advances in large language models have spurred research into enhancing their reasoning capabilities, particularly in math-rich STEM documents. |
| Approach: | They propose a benchmark dataset to evaluate LLMs’ reasoning abilities on math symbols within contextual scientific text. |
| Outcome: | The proposed dataset demonstrates that state-of-the-art LLMs achieve an average accuracy of 20-60% under in-context learning and 50-60% with fine-tuning, highlighting a substantial gap in their ability to classify mathematical symbols. |
Can NLI Models Verify QA Systems’ Predictions? (2021.findings-emnlp)
Copied to clipboard
| Challenge: | Recent question answering systems perform well on benchmark datasets, but are not always well-calibrated to spot spurious answers under distribution shifts. |
| Approach: | They propose to use natural language inference to verify whether answers are correct . they leverage large pre-trained models and recent prior datasets to construct powerful question conversion and decontextualization modules. |
| Outcome: | The proposed approach improves the confidence estimation of a QA model across different domains, evaluated in a selective QA setting. |
Tree-Based Representation and Generation of Natural and Mathematical Language (2023.acl-long)
Copied to clipboard
| Challenge: | Existing models for generating and modeling mathematical language are limited . existing models for modeling and generating mathematical language simply treat mathematical expressions as text . |
| Approach: | They propose to combine mathematical expressions and text-based models to generate mathematically valid expressions. |
| Outcome: | The proposed model outperforms baselines on mathematical expression generation tasks. |
Let’s Reason Formally: Natural-Formal Hybrid Reasoning Enhances LLM’s Math Capability (2025.emnlp-main)
Copied to clipboard
| Challenge: | Recent work has focused on improving the mathematical reasoning capabilities of Large Language Models (LLMs). |
| Approach: | They propose an end-to-end framework to integrate FL into NL math reasoning . they propose a problem alignment method that reformulates QA and existence problems . |
| Outcome: | The proposed framework achieves 89.80% and 84.34% accuracy rates on the MATH-500 and the AMC benchmarks. |
Complex Reasoning in Natural Language (2023.acl-tutorials)
Copied to clipboard
| Challenge: | Recent research shows that pretrained language models are often brittle for complex reasoning tasks. |
| Approach: | They propose to use pre-trained language models to teach machines to reason over texts . they will review recent promising approaches to tackling complex reasoning tasks . |
| Outcome: | This tutorial reviews promising approaches to complex reasoning tasks . it reviews the methods that can be used to augment models with robustness . |
LLMs for Mathematical Modeling: Towards Bridging the Gap between Natural and Mathematical Languages (2025.findings-naacl)
Copied to clipboard
| Challenge: | Large Language Models (LLMs) have demonstrated strong performance across various natural language processing tasks, but their proficiency in mathematical reasoning remains a key challenge. |
| Approach: | They propose a process-oriented framework to evaluate LLMs' ability to construct mathematical models, using solvers to compare outputs with ground truth. |
| Outcome: | The proposed framework evaluates LLMs' ability to construct mathematical models, using solvers to compare outputs with ground truth. |