Papers by Yuan Peiyue
OCR or Not? Rethinking Document Information Extraction in the MLLMs Era with Real-World Large-Scale Datasets (2026.eacl-industry)
Copied to clipboard
| Challenge: | Multimodal Large Language Models (MLLMs) are used for document information extraction, but their impact on document information processing remains unclear. |
| Approach: | They propose an automated hierarchical error analysis framework that leverages large language models to diagnose errors systematically. |
| Outcome: | The proposed framework can achieve comparable performance to OCR-enhanced approaches. |