Papers with RLAE_PPO

    1 papers
    RLAE: Reinforcement Learning-Assisted Ensemble for LLMs (2025.emnlp-main)

    Copied to clipboard

    Challenge: Existing ensemble methods for ensembling large language models rely on fixed weighting strategies that fail to adapt to dynamic, context-dependent characteristics of LLMs.
    Approach: They propose a framework that reformulates LLM ensemble through a Markov Decision Process.
    Outcome: The proposed framework outperforms existing methods by 3.3% on a diverse set of tasks while achieving lower time latency.

    What is GenGO?

    GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

    Information

    About
    Limitations