Papers by Weilei Wang
CRAB: A Benchmark for Evaluating Curation of Retrieval-Augmented LLMs in Biomedicine (2025.emnlp-industry)
Copied to clipboard
| Challenge: | Recent development in Retrieval-Augmented Large Language Models (LLMs) have shown great promise in biomedical applications. |
| Approach: | They propose a multilingual benchmark to evaluate retrieval-augmented large language models' curation ability. |
| Outcome: | The proposed benchmark is available in English, French, German and Chinese. |