Physical and Computational Approaches Towards More Scalable Molecular Systems

dc.contributor.advisorNivala, Jeffrey
dc.contributor.advisorCeze, Luis H
dc.contributor.authorStephenson, Ashley Paige
dc.date.accessioned2026-08-11T19:26:55Z
dc.date.issued2026-08-11
dc.date.submitted2026
dc.descriptionThesis (Ph.D.)--University of Washington, 2026
dc.description.abstractThe ability to understand and engineer biological systems depends on efficiently manipulating and interpreting molecular information. Design-Build-Test-Learn (DBTL) cycles remain limited by bottlenecks in both experimental workflows and computational analysis. This thesis addresses these complementary challenges through both physical laboratory automation and machine learning methods for interpreting nanopore sequencing signals from chemically diverse biopolymers. First, I present work on digital microfluidic automation using the PurpleDrop platform, enabling programmable droplet manipulation for applications including DNA data storage and aptamer discovery. This work demonstrates the potential of integrated hardware-software systems for automated experimentation while highlighting challenges in reliability and system integration. The primary focus of this thesis develops structure-informed machine learning models for nanopore sequencing of expanded molecular alphabets. Xenonucleic acids (XNAs), synthetic nucleotides with modified bases, sugars, or backbones, hold promise for therapeutics, diagnostics, and information storage but lack robust sequencing methods. Existing nanopore approaches require extensive experimental calibration across many sequence contexts. I show that models incorporating molecular structure and chemical features can predict nanopore ionic current signatures for unseen XNA bases, substantially reducing experimental requirements. Structure-informed models outperform sequence-only approaches in few-shot and zero-shot settings and enable practical XNA sequencing with limited calibration data. Finally, I extend this framework to nanopore protein sequencing for detection of post-translational modifications such as phosphorylation. Despite the greater complexity of proteins, similar modeling approaches enable phosphorylation detection, suggesting broader applicability of structure-based methods for complex biopolymer sequencing. Together, these contributions reduce experimental burden through both physical automation and computational prediction, enabling more scalable DBTL cycles for increasingly complex molecular systems.
dc.embargo.termsOpen Access
dc.format.mimetypeapplication/pdf
dc.identifier.otherStephenson_washington_0250E_29959.pdf
dc.identifier.urihttps://hdl.handle.net/1773/57256
dc.language.isoen_US
dc.rightsnone
dc.subjectAutomation
dc.subjectMachine learning
dc.subjectMicrofluidics
dc.subjectMolecular biology
dc.subjectNanopore sequencing
dc.subjectSynthetic biology
dc.subjectMolecular biology
dc.subjectComputer science
dc.subject.otherComputer science and engineering
dc.titlePhysical and Computational Approaches Towards More Scalable Molecular Systems
dc.typeThesis

Files

Original bundle

Now showing 1 - 1 of 1
Loading...
Thumbnail Image
Name:
Stephenson_washington_0250E_29959.pdf
Size:
20.05 MB
Format:
Adobe Portable Document Format