Papers with TikTalkCoref
Multimodal Coreference Resolution for Chinese Social Media Dialogues: Dataset and Benchmark Approach (2025.acl-long)
Copied to clipboard
| Challenge: | Multimodal coreference resolution (MCR) aims to identify mentions referring to the same entity across different modalities, such as text and visuals. |
| Approach: | They propose a Chinese multimodal coreference dataset based on Douyin short-video platform to help researchers understand multimodal content. |
| Outcome: | The proposed dataset pairs short videos with corresponding textual dialogues from user comments and includes manually annotated coreference clusters for person mentions in the text and the coreferential person head regions in the corresponding video frames. |