Investigating Generalization of Unlike Coordination in Language Models via Filtered-Corpus Training

dc.contributor.advisorSteinert-Threlkeld, Shane
dc.contributor.authorLuo, Jiamu
dc.date.accessioned2026-08-11T19:32:08Z
dc.date.issued2026-08-11
dc.date.submitted2026
dc.descriptionThesis (Master's)--University of Washington, 2026
dc.description.abstractA long-standing debate in theoretical linguistics concerns the nature of coordination. A common view holds that only same-category constituents can be conjoined, which has been challenged by the many grammatical cases of unlike coordination found in natural language. This thesis investigates how language models (LMs) learn and generalize the linguistic phenomenon of coordination by Filtered-Corpus Training, which creates an environment where models have only been exposed to alike coordination during training. We evaluate and compare these models to counterparts trained on an unfiltered corpus. Our results suggest unlike coordination is not a general exception to LMs and can be learned only with indirect information, although direct exposure may still be needed for the more challenging cases. They further indicate that LMs process unlike coordination by treating the conjoined elements as belonging to similar structural categories or through a mechanism akin to deletion, both of which appear learnable from exposure to alike coordination alone. This work contributes to the growing understanding of how language models internally represent linguistic structure, while also adding to the broader debate on coordination by showing that how models generalize and process unlike coordination without direct exposure.
dc.embargo.termsOpen Access
dc.format.mimetypeapplication/pdf
dc.identifier.otherLuo_washington_0250O_29571.pdf
dc.identifier.urihttps://hdl.handle.net/1773/57434
dc.language.isoen_US
dc.rightsCC BY-NC-ND
dc.subjectCoordination
dc.subjectFiltered Corpus Training
dc.subjectInductive Biases
dc.subjectLinguistic Generalization
dc.subjectPoverty of the Stimulus
dc.subjectLinguistics
dc.subjectArtificial intelligence
dc.subject.otherLinguistics
dc.titleInvestigating Generalization of Unlike Coordination in Language Models via Filtered-Corpus Training
dc.typeThesis

Files

Original bundle

Now showing 1 - 1 of 1
Loading...
Thumbnail Image
Name:
Luo_washington_0250O_29571.pdf
Size:
1.03 MB
Format:
Adobe Portable Document Format

Collections