Papers by Yafeng Chen

4 papers
Enriching Word Embeddings with Domain Knowledge for Readability Assessment (C18-1)

Copied to clipboard

Challenge: Existing word embedding models focus on syntactic or semantic relations of words, while ignoring reading difficulty.
Approach: They propose a method which learns the word embedding for readability assessment . they extract the knowledge on word-level difficulty from three perspectives to construct a knowledge graph .
Outcome: The proposed method is effective and potential, the authors show . they use the knowledge-enriched word embedding model on English and Chinese datasets .
Integrating Audio, Visual, and Semantic Information for Enhanced Multimodal Speaker Diarization on Multi-party Conversation (2025.acl-long)

Copied to clipboard

Challenge: Mainstream speaker diarization systems rely only on acoustic information, making it challenging in complex aural environments.
Approach: They propose a multimodal approach that integrates audio, visual, and semantic cues to enhance speaker diarization.
Outcome: The proposed approach outperforms state-of-the-art methods on multi-party conversations . it integrates audio-visual-semantic cues into the clustering process for acoustic speaker embeddings .
Exploring Speaker-Related Information in Spoken Language Understanding for Better Speaker Diarization (2023.findings-acl)

Copied to clipboard

Challenge: Current speaker diarization systems consider only acoustic information, resulting in performance degradation when encountering adverse acustic environment.
Approach: They propose methods to extract speaker-related information from conversational semantics in multi-party meetings.
Outcome: The proposed method improves on AISHELL-4 and AliMeeting datasets on speakers diarization and speaker-turn detection.
Multi-Prompting Decoder Helps Better Language Understanding (2025.findings-acl)

Copied to clipboard

Challenge: Existing methods to adapt Pre-trained Language Models to downstream tasks are limited by their inference APIs.
Approach: They propose a multi-prompting decoding framework that query PLMs with multiple prompts . they propose to query Plms with optimal transport for hidden states and calibrated decoding for class scores .
Outcome: The proposed framework achieves state-of-the-art results on multiple natural language understanding datasets under the few-shot setting.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations