Measuring the perceptual availability of phonological features during language acquisition using unsupervised binary stochastic autoencoders (N19-1)
Copied to clipboard
| Challenge: | Xitsonga and English are typologically unrelated languages . phonological features are not directly observed by humans . |
| Approach: | They deploy binary stochastic neural autoencoder networks as models of infant language learning in two typologically unrelated languages. |
| Outcome: | The proposed model is well represented in both languages, while others are less so. |
Similar Papers
Large Language Models Share Representations of Latent Grammatical Concepts Across Typologically Diverse Languages (2025.naacl-long)
Copied to clipboard
| Challenge: | In the brains of human bilinguals, syntax processing may occur in similar regions for their first and second language, depending on factors like when the second language was learned and language proficiency. |
| Approach: | They propose to use sparse autoencoders to train Llama-3-8B and Aya-23-8B models to train multilingual models that share morphsyntactic representations of grammatical concepts. |
| Outcome: | The proposed model can predict plural verbs in different languages by activating the same plural feature. |
From Phonology to Syntax: Unsupervised Linguistic Typology at Different Levels with Language Embeddings (N18-1)
Copied to clipboard
| Challenge: | linguistic typology is the classification of languages according to their linguistic properties. |
| Approach: | They learn distributed language representations which can be used to predict typological properties on a massively multilingual scale. |
| Outcome: | The proposed model can predict typological properties on a massively multilingual scale. |
Exploring How Generative Adversarial Networks Learn Phonological Representations (2023.acl-long)
Copied to clipboard
| Challenge: | Recent studies in natural language processing (NLP) have demonstrated two generic trends: neural networks dominate language-specific machine learning models; the interpretability of these models is limited that the language representation they learned might not align to human language. |
| Approach: | They propose to use a phonological feature-learning architecture to encode contrastive and non-contrastive nasality in French and English vowels. |
| Outcome: | The proposed architecture encodes contrastive and non-contrastive nasality in French and English vowels. |
Transformer-based Speech Model Learns Well as Infants and Encodes Abstractions through Exemplars in the Poverty of the Stimulus Environment (2025.coling-main)
Copied to clipboard
| Challenge: | Existing theories of language learning for infants are inadequate, according to Chomsky . infants learn language in impoverished environments, according a new study . |
| Approach: | They designed a series of tasks, scenarios, and metrics to simulate the POS . they found that the emerging speech model wav2vec2.0 can learn well in noisy Mandarin environments. |
| Outcome: | The proposed model can learn in noisy and sparse Mandarin environments. |
Emergent morpho-phonological representations in self-supervised speech models (2025.emnlp-main)
Copied to clipboard
| Challenge: | a recent study shows that self-supervised speech models do not represent phonological and morphological phenomena in frequent English noun and verb inflections. |
| Approach: | They study how S3Ms represent phonological and morphological phenomena in English . they propose alternative representational strategies that may support human spoken word recognition . |
| Outcome: | a new study shows that S3M models can represent phonological and morphological phenomena in English . the models can be trained to recognize spoken words in naturalistic, noisy environments . |
Layer-wise Minimal Pair Probing Reveals Contextual Grammatical-Conceptual Hierarchy in Speech Representations (2025.emnlp-main)
Copied to clipboard
| Challenge: | a recent study evaluated the extent to which SLMs encode nuanced syntactic and conceptual features . acoustic and phonetic features are shallow, but the extent of nuance is unclear . |
| Approach: | a new study evaluates contextual syntactic and semantic features in transformer-based speech language models . authors compare SLMs to linguistic competence assessments for large language models. |
| Outcome: | a new study compares SLMs with linguistic competence assessments to assess speech recognition and understanding . the results show that SLM models encode grammatical features more robustly than conceptual ones . |
Small Language Models Also Work With Small Vocabularies: Probing the Linguistic Abilities of Grapheme- and Phoneme-Based Baby Llamas (2025.coling-main)
Copied to clipboard
| Challenge: | Existing studies on LMs have focused on linguistic generalizations and representations from developmentally plausible data. |
| Approach: | They propose to use phoneme- and grapheme-based language models to learn linguistic units at and below the word level. |
| Outcome: | The proposed models can achieve strong performance on syntactic and novel benchmarks and match grapheme-based models in standard tasks and novel evaluations. |
Can Language Models Learn Typologically Implausible Languages? (2026.tacl-1)
Copied to clipboard
| Challenge: | Language models provide a naturalistic framework for studying artificial language learning . authors: typological universals and tendencies are thought to be caused by a learning bias . |
| Approach: | They propose to train LMs on highly naturalistic counterfactual versions of English and Japanese . they show that LM learn subtly implausible languages more slowly . |
| Outcome: | The proposed language models learn subtly implausible languages more slowly compared to human models . the findings suggest that LMs exhibit typologically aligned learning preferences . |
Cross-Lingual Generalization and Compression: From Language-Specific to Shared Neurons (2025.acl-long)
Copied to clipboard
| Challenge: | Existing evidence suggests that multilingual language models can transfer knowledge across languages without explicit cross-lingual supervision. |
| Approach: | They analyze the parameter spaces of three multilingual language models to examine their representations . they find that models evolve from language-specific representations to more specialized layer functions . |
| Outcome: | The proposed model can generate coherent English text, rather than spanish text, and it can generate generalized representations, the authors show. |
What Do Neural Speech Models Know About Phonology? Evidence from Structured Phoneme Confusions (2026.findings-acl)
Copied to clipboard
| Challenge: | acoustic and phonological models of speech recognition are often limited to the phoneme level . a recent study has shown that phoneme confusions are strongly structured in phonology space . |
| Approach: | They adopt a featural representation of phonemes grounded in phonological theory which models speech sounds as structured bundles of distinctive articulatory and acoustic properties. |
| Outcome: | The proposed model allows us to analyse phoneme confusions at a finer granularity and to investigate whether certain phonological features are more vulnerable than others. |