Papers by Samuel Amouyal

2 papers
Large Language Models for Psycholinguistic Plausibility Pretesting (2024.findings-eacl)

Copied to clipboard

Challenge: Psycholinguists typically use language models to create controlled materials . plausibility judgments are often based on coarse-grained judgements, but fine-grounded ones do not .
Approach: They investigate whether Language Models can be used to generate plausibility judgments . they find that plausible judgements from LMs are highly related to human judgements - whereas other LM models are not .
Outcome: The proposed language models can generate plausibility judgments from human evaluators . the proposed models do not provide satisfactory discriminative power .
AssistantBench: Can Web Agents Solve Realistic and Time-Consuming Tasks? (2024.emnlp-main)

Copied to clipboard

Challenge: Current language models and retrieval-augmented LMs are limited in their ability to perform tasks on the web.
Approach: They propose a benchmark to evaluate language agents built on top of language models . they propose 'AssistantBench' which includes 214 tasks that can be automatically evaluated .
Outcome: The proposed agent outperforms existing agents in a new benchmark for language agents on the web.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations