Papers by Zifeng Zhao

    2 papers
    SciAssess: Benchmarking LLM Proficiency in Scientific Literature Analysis (2025.findings-naacl)

    Copied to clipboard

    Challenge: Existing benchmarks fail to adequately evaluate the proficiency of Large Language Models (LLMs) Existing standards do not cover the skills needed to evaluate LLMs in scientific literature analysis.
    Approach: They propose a benchmark to evaluate the proficiency of large language models in scientific literature analysis.
    Outcome: SciAssess evaluates 11 LLMs on multiple tasks across scientific fields.
    ProtoCycle: Reflective Tool-Augmented Planning for Text-Guided Protein Design (2026.findings-acl)

    Copied to clipboard

    Challenge: Recent deep generative models have already shown encouraging * Equal contribution.
    Approach: They propose to use generic instruction-tuned LLMs as direct text-to-sequence generators to achieve this goal.
    Outcome: Recent studies show that reflection improves sequence quality and alignment while maintaining competitive foldability.

    What is GenGO?

    GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

    Information

    About
    Limitations