A data-driven sequencer that unveils latent “codons” in synthetic copolymers

13Citations
Citations of this article
13Readers
Mendeley users who have this article in their library.

Abstract

The recent emergence of sequence engineering in synthetic copolymers has been innovating polymer materials, where short sequences, hereinafter called “codons” using an analogy from nucleotide triads, play key roles in expressing functions. However, the codon compositions cannot be experimentally determined owing to the lack of efficient sequencing methods, hindering the integration of experiments and theories. Herein, we propose a polymer sequencer based on mass spectrometry of pyrolyzed oligomeric fragments. Despite the random fragmentation along copolymer main-chains, the characteristic fragment patterns of the codons are identified and quantified via unsupervised learning of a spectral dataset of random copolymers. The codon complexities increase with their length and monomer component number. Our data-driven approach accommodates the increasing complexities by expanding the dataset; the codon compositions of binary triads, binary pentads and ternary triads are quantifiable with small datasets (N < 100). The sequencer allows describing copolymers with their codon compositions/distributions, facilitating sequence engineering toward innovative polymer materials.

Cite

CITATION STYLE

APA

Hibi, Y., Uesaka, S., & Naito, M. (2023). A data-driven sequencer that unveils latent “codons” in synthetic copolymers. Chemical Science, 14(21), 5619–5626. https://doi.org/10.1039/d2sc06974a

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free