Papers by Yuan Peiyue

1 papers
OCR or Not? Rethinking Document Information Extraction in the MLLMs Era with Real-World Large-Scale Datasets (2026.eacl-industry)

Copied to clipboard

Challenge: Multimodal Large Language Models (MLLMs) are used for document information extraction, but their impact on document information processing remains unclear.
Approach: They propose an automated hierarchical error analysis framework that leverages large language models to diagnose errors systematically.
Outcome: The proposed framework can achieve comparable performance to OCR-enhanced approaches.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations