Papers by Prisha Samdarshi

1 papers
Connecting the Dots: Evaluating Abstract Reasoning Capabilities of LLMs Using the New York Times Connections Word Game (2024.emnlp-main)

Copied to clipboard

Challenge: We evaluate the performance of large language models (LLMs) against expert and novice human players.
Approach: They propose to use the New York Times Connections game as a test bed to evaluate the abstract reasoning capabilities of large language models (LLMs) they propose to test the ability of large-language models to be able to cluster and categorize words using semantic relations.
Outcome: The proposed game is a test bed for evaluating abstract reasoning capabilities in humans and AI systems.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations