Papers by Abrar Anwar
Which One? Leveraging Context Between Objects and Multiple Views for Language Grounding (2024.naacl-long)
Copied to clipboard
| Challenge: | Existing methods for identifying object referents of language expressions consider target and distractor objects independently and pool multiple views before grounding. |
| Approach: | They propose a model that selects an object referent based on language that distinguishes between two similar objects and a multi-view approach to grounding in context model which reduces the relative error by 12.9% . |
| Outcome: | The proposed model improves on the SNARE object reference task with a relative error reduction of 12.9% and an absolute improvement of 2.7%. |