Papers by Ryota Tanaka

1 papers
How Well Do Vision Models Encode Diagram Attributes? (2024.acl-srw)

Copied to clipboard

Challenge: Experimental results show vision models struggle to identify diagram attributes such as node colors and shapes, along with edge colors and connection patterns.
Approach: They evaluated vision models and retrieving diagrams using text queries to determine how well they recognize diagram attributes and edge connection patterns.
Outcome: The models can recognize node colors, shapes, and edge colors, but struggle to identify differences in edge connection patterns that play a pivotal role in the semantics of diagrams.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations