Papers by Martin Masek
Spatially-Aware Speaker for Vision-and-Language Navigation Instruction Generation (2024.acl-long)
Copied to clipboard
| Challenge: | Existing speaker models learn strategies to evade evaluation metrics and obtain higher scores even for low-quality sentences. |
| Approach: | They propose a speaker-based instruction generator that utilises both structural and semantic knowledge of the environment to produce richer instructions. |
| Outcome: | The proposed model outperforms existing models and is evaluated using standard metrics. |