Papers by Nathaniel Blanchard

4 papers
Common Ground Tracking in Multimodal Dialogue (2024.lrec-main)

Copied to clipboard

Challenge: In dialogue modeling, there is considerable attention on “dialogue state tracking” (DST) but “common ground tracking” identifies the shared belief space held by all participants in a task-oriented dialogue: the task-relevant propositions all participants accept as true.
Approach: They propose a method for automatically identifying the current set of shared beliefs and ”questions under discussion” of a group with a shared goal.
Outcome: The proposed method predicts moves toward building common ground relative to ground truth in a multimodal interaction with an AI.
Multimodal Cross-Document Event Coreference Resolution Using Linear Semantic Transfer and Mixed-Modality Ensembles (2024.lrec-main)

Copied to clipboard

Challenge: Existing methods for cross-document coreference resolution do not provide images for all mentions of events.
Approach: They propose a multimodal cross-document event coreference resolution method that integrates visual and textual cues with a simple linear map between vision and language models.
Outcome: The proposed method improves on a popular ECB+ and AIDA datasets.
“Any Other Thoughts, Hedgehog?” Linking Deliberation Chains in Collaborative Dialogues (2024.findings-emnlp)

Copied to clipboard

Challenge: Recent advances in generative AI have raised the possibility of systems that follow and interact with multiparty dialogues.
Approach: They propose a graph-based framework for probing questions in collaborative dialogues that models causal relations between probing and causal utterances and the links between them.
Outcome: The proposed framework compares to baselines and stronger coreference approaches and establishes a standard of performance in this novel task.
The VoxWorld Platform for Multimodal Embodied Agents (2022.lrec-1)

Copied to clipboard

Challenge: a retrospective of the VoxWorld platform is presented . it is a platform for rapidly building and deploying embodied agents with contextual and situational awareness.
Approach: They present a retrospective on the development of the VoxWorld platform . they focus on three different agent implementations and the functionality needed to accommodate them .
Outcome: The VoxWorld platform has evolved from a theoretical model to a platform capable of multimodal interaction and hybrid reasoning.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations