Papers by Sungkyung Kim

1 papers
Towards Efficient Visual-Language Alignment of the Q-Former for Visual Reasoning Tasks (2024.findings-emnlp)

Copied to clipboard

Challenge: Pre-trained large language models can be fine-tuned with instruction tuning to align the model responses with human intentions.
Approach: They investigate the effectiveness of parameter efficient fine-tuning (PEFT) of the Q-Former with visual reasoning benchmarks ScienceQA and IconQA.
Outcome: The proposed model achieves comparable performance to full fine-tuning using under 2% of the trainable parameters.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations