Papers with Room-for-Room

2 papers
Stay on the Path: Instruction Fidelity in Vision-and-Language Navigation (P19-1)

Copied to clipboard

Challenge: Existing metrics for vision-and-language navigation focus on goal completion rather than the sequence of actions corresponding to the instructions.
Approach: They propose to use a room-to-room dataset to measure the length of instruction followed by agents.
Outcome: The proposed metric outperforms existing metrics for Room-to-Room tasks because it is direct-to goal shortest.
Masked Path Modeling for Vision-and-Language Navigation (2023.findings-emnlp)

Copied to clipboard

Challenge: A major challenge in vision-and-language navigation is the limited available training data, which hinders the models’ ability to generalize effectively.
Approach: They propose a masked path modeling objective that pretrains an agent using self-collected data for subsequent navigation tasks.
Outcome: The proposed model pretrains an agent using self-collected data for subsequent navigation tasks eliminating the need for external tools.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations