Papers by Yefan Zhou

2 papers
Model Balancing Helps Low-data Training and Fine-tuning (2024.emnlp-main)

Copied to clipboard

Challenge: Recent advances in foundation models have emphasized the need to align pre-trained models with specialized domains using small, curated datasets.
Approach: They propose a layer-wise learning rate scheduler that balances training quality across layers . they adapt it to a curated dataset to achieve alignment with specialized domains .
Outcome: The proposed model shows that it can be used to balance training quality across layers and improve low-data training and fine-tuning for both NLP and SciML tasks.
AlphaLoRA: Assigning LoRA Experts Based on Layer Training Quality (2024.emnlp-main)

Copied to clipboard

Challenge: Recent studies combine LoRA with Mixture-of-Experts (MoE) to improve performance in Large Language Models.
Approach: They propose a method to combine LoRA and Mixture-of-Experts (MoE) to improve performance in Large Language Models.
Outcome: The proposed method reduces redundancy in LoRA experts within the MoE architecture, and improves training quality across layers.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations