Papers by Angela Lin

2 papers
A Chinese Dataset for Evaluating the Safeguards in Large Language Models (2024.findings-acl)

Copied to clipboard

Challenge: a recent study has shown that large language models can produce harmful responses, exposing users to unexpected risks.
Approach: They propose a dataset for the safety evaluation of Chinese LLMs in Mandarin Chinese . they extend the dataset to better identify false negative and false positive examples .
Outcome: The proposed dataset is for the safety evaluation of Chinese LLMs, and is based on a Chinese dataset.
A Recipe for Creating Multimodal Aligned Datasets for Sequential Tasks (2020.acl-main)

Copied to clipboard

Challenge: a web-based algorithm can be used to align instructions for different tasks . video instructions can be noisy and contain far more information than textual instructions.
Approach: They propose an algorithm that learns pairwise alignments between different recipes . they then use a graph algorithm to derive a joint alignment between multiple video and text recipes based on the same recipe.
Outcome: The proposed algorithm learns pairwise alignments between different recipes for the same dish.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations