Papers by Shuhua Yang
HTMuon: Improving Muon via Heavy-Tailed Spectral Correction (2026.findings-acl)
Copied to clipboard
| Challenge: | Muon’s orthogonalized update rule suppresses the emergence of heavy-tailed weight spectra and over-emphasizes the training along noise-dominated directions. |
| Approach: | They propose to preserve Muon's ability to capture parameter interdependencies while producing heavier-tailed updates and inducing heavier-tail weight spectra. |
| Outcome: | The proposed algorithm suppresses the emergence of heavy-tailed weight spectra and over-emphasizes training along noise-dominated directions. |
Query-Efficient Agentic Graph Extraction Attacks on GraphRAG Systems (2026.acl-long)
Copied to clipboard
| Challenge: | Existing attacks exploit leakage of retrieved subgraphs, leaving the security implications of structured knowledge representations unexplored. |
| Approach: | They propose a framework that leverages a novelty-guided exploration–exploitation strategy and external graph memory modules to extract a latent entity–relation graph. |
| Outcome: | The proposed framework outperforms baselines on medical, agriculture, and literary datasets under identical query budgets while maintaining high precision. |
A Unified Supervised and Unsupervised Dialogue Topic Segmentation Framework Based on Utterance Pair Modeling (2025.naacl-long)
Copied to clipboard
| Challenge: | Unsupervised methods for dialogue topic segmentation are difficult to surpass due to short sentences, serious references and non-standard language. |
| Approach: | They propose a method to divide a dialogue into different topic paragraphs to better understand its structure and content. |
| Outcome: | The proposed method achieves the best results on multiple benchmark datasets across different scenarios. |