Papers by Hongkuan Zhang
Cross-Modal Similarity-Based Curriculum Learning for Image Captioning (2022.emnlp-main)
Copied to clipboard
| Challenge: | Existing image captioning approaches treat image-caption pairs indistinctly without considering the differences in their learning difficulties. |
| Approach: | They propose a pretrained vision–language model that measures cross-modal similarity and a model that uses cross-module similarity to measure the difficulty of captioning. |
| Outcome: | The proposed model achieves superior performance and competitive convergence speed to baselines without incurring additional training costs. |
Development of a Medical Incident Report Corpus with Intention and Factuality Annotation (2020.lrec-1)
Copied to clipboard
| Challenge: | Medical incident reports are documents that record what happened in a medical incident. |
| Approach: | They propose to annotate medical incident reports with annotations of intention and factuality and medication entities and their relations. |
| Outcome: | The proposed method combines the definition of medication entities and the method to annotate the relations between entities and extracts important information from the unstructured part. |