Papers by Dong-Kyum Kim
How Training Data Shapes the Use of Parametric and In-Context Knowledge in Language Models (2026.acl-long)
Copied to clipboard
| Challenge: | Large language models leverage parametric and in-context knowledge in training . however, when these sources conflict, models arbitrate based on their internal confidence . |
| Approach: | They conduct controlled experiments using synthetic corpora to identify data properties that shape knowledge utilization. |
| Outcome: | The results show that the robust use of both knowledge sources is an emergent property . the results provide guidance for designing training data that supports the reliability of parametric and in-context knowledge in language models. |