Pseudorandom Vs. Random Polymers - How to Improve the Efficiency of Lithography-based Synthesis
Author Affiliations
VP Bioinformatics, Caris Life Sciences, Phoenix, Arizona, United States
Corresponding Author
Phillip Stafford, VP Bioinformatics, Caris Life Sciences, Arizona, United States
Citation
© 2018 Stafford P. This is an openaccess article distributed under the terms of the
Creative Commons Attribution 4.0 international
License, which permits unrestricted use,
distribution and reproduction in any medium,
provided the original author and source are
credited.
Abstract
Keywords
in situ peptide synthesis; Photolithography; Random peptide; Immunosignature
Introduction
In situ peptide synthesis
There are many ways to manufacture peptide microarrays. For a
review see Gao et al. [27,28]. Light-directed chemical synthesis has
been quite successful, dating back to 1995 (Fodor et al. [29], US Patent
5405783A). Some synthesis methods, direct light on a photoacid [16]
or photobase [30] generator which removes amino acid protecting
groups. Other methods use photoactivatable amino acids [29]. Some
methods use shadow masks [16], some use digital micromirrors [31],
some utilize a peptide ‘toner’ and a 20 ‘color’ laser printer [32], while
still others use electrical circuits to alter the pH in a microwell [33].
Each of these synthesis methods relies on the systematic addition of
amino acids to a growing peptide. For random peptide libraries, all of
these methods can take advantage of reduced synthetic steps using
pseudorandom peptides.
One of the highest throughput methods for the production of
peptide microarrays is shadow-mask lithography [34]. Shadowmasks are pre-made so they are not as flexible as programmable
light-directed arrays, and they require substantial up-front costs.
However, by implementing high-resolution steppers, automated
aligners, synthesizers and large wafers, mask-based peptide
microarrays can be very inexpensive, high density, high precision
and high throughput. These characteristics are important for high
volume/low costproduction.
Random vs. pseudorandom libraries
Many immunological experiments make use of random sequence
peptides [35]. Phage display uses random sequence libraries [36]
and has been used to identify therapeutics [37], epitopes [38], even
enzyme modulators [39]. Panning large libraries can identify very
specific biomarkers, but these methods are neither cheap nor high
throughput. Immunosignatures use a fixed library of a given size
for any sample. A diverse surface of 3D shapes produced by random
sequence peptides has been very successful in many applications
[11,13,21,40-42]. Surprisingly, 10,000 peptides were sufficient to
distinguish over 15 different diseases simultaneously, with >95%
accuracy [14]. However, 10,000 peptides likely had a finite level of
specificity and sensitivity. It was decided to create at least 300,000
peptides for the next immunosignature iteration. When M=AA*L,
synthesis of 300,000 17mers would take over 2 weeks of continuous
synthesis. Altering the library from random to pseudorandom could
drastically reduce manufacturing time.
How in-situ peptide synthesis is affected by amino acid
selection
The simplest algorithm for the number of synthesis steps (or
photomasks) required to produce a peptide of length L is M (number
of masks)=AA (number of amino acids used) * L (length). For a 17mer
using 20 amino acids, 20*17=340 synthetic steps are required. In
mask-based peptide photolithography, 340 steps would take over
2 weeks to complete. For immunosignature peptides, M can be less
than AA*L, but the resulting peptides become pseudorandom. Each
mask costs several thousand dollars. Each synthesis step takes
approximately 15 minutes. It is beneficial to reduce M. Figure 1
illustrates the bias imposed by forcing M to be less than AA*L.
Experimental design
We asked two questions about reducing synthesis steps. First,
how do the chemical characteristics of the peptide library change
as mask numbers are reduced? Second, how does reduced diversity
alter the outcome of a biochemical assay? The first question was
addressed using predicted characteristics of hypothetical peptides.
For the second question, an immunosignature of monoclonal
antibodies and an immunosignature of human sera are used to
measure performance. We first established the starting parameters.
Kuznetsov [43] noted that some amino acids are preferentially used
in immunosignatures [43]. To maintain consistency with the actual
immunosignature arrays that will be manufactured, we eliminated
Methionine, Threonine and Isoleucine because they appear rarely in
peptides that make up diagnostic immunosignatures [43]. Cysteine can form disulfide bonds under aqueous conditions, and was
removed for that reason. Thus our virtual 17mer peptide library used
16 amino acids (no C, T, M, or I). We created two control sets: M=272
uses 16 amino acids and no other mask restrictions. M=323 leaves
out only Cysteine. The M=323 case matches exactly the existing
10,000 peptide spotted peptide library that has been used in several
early publications (NCBI GEO accession GPL17600). We also created
intermediate test sets where M=35, M=70, and M=140.
The first analysis was completed using 100,000 virtual peptides.
For each library, an algorithm (see Supplemental Information) creates
M masks with 100,000 features and ‘synthesizes’ peptides using
these virtual masks. The resulting library is analyzed for chemical
properties using ProtParam ((http://web.expasy.org/protparam/).
The second analysis involved synthesizing a library of actual
peptides corresponding to the same mask restrictions. 384 peptides
were selected at random from each 100,000 virtual peptide libraries
and sent to Sigma Genosys (St. Louis, MO) for synthesis. We already
had several M=323 10,000 peptide microarrays made, so the human
sera experiments for M=323 were done by randomly picking 384
peptides from the 10,000 existing peptides to simulate a 384-peptide
library. Each of these 384-peptide sets (except M=323) were
tested using 51 different commercial monoclonals (Supplemental
Information Table 1). We previously published data on monoclonals
using the 10,000 peptide microarray [7]. Human sera from 8 controls
(healthy) humans and 8 patients diagnosed with Coccidiomycoses
[11] were used for all libraries including M=323.
Materials and Methods
Table 1
Figure 3: Display of unique 2mer, 3mer, 4mer, 5mer and 6mers
contained within 100,000 peptides using the M=35, M=70, M=140,
M=272 and M=323 mask restriction sets.
The Y-axis is the proportion of unique nmers, and the X-axis is
increasing length of subsequences.
Conclusions
Acknowledgement
There were no funding agencies for this work. Data are
freely available at the Gene Expression Omnibus at NCBI. Platforms:
GPL17600, GPL17490, GPL17679, Experimental Series: GSE49217,
GSE50044, GSE50045.
References
- Buus S, Rockberg J, Forsström B, Nilsson P, Uhlen M, et al. Highresolution Mapping of Linear Antibody Epitopes Using Ultrahighdensity Peptide Microarrays. Mol Cell Proteomics. 2012
Dec;11(12):1790-800.
- Domenyuk V, Loskutov A, Johnston SA, Diehnelt CW. A Technology
for Developing Synbodies with Antibacterial Activity. PLoS ONE.
2013 Jan;8:e54162.
- Boltz KW, Gonzalez-Moa MJ, Stafford P, Johnston SA, Svarovsky
SA. Peptide microarrays for carbohydrate recognition. Analyst.
2009 Apr;134(4):650-652.
- Sinzinger MD, Chung YD, Adjobo-Hermans MJW, Brock R.
A microarray-based approach to evaluate the functional
significance of protein-binding motifs. Analytical Anal Bioanal
Chem. 2016 Feb;408:3177–3184.
- Houseman BT, Huh JH, Kron SJ, Mrksich M. Peptide chips for the
quantitative evaluation of protein kinase activity. Nat Biotech.
2002 Mar;20:270-274.
- Fu J. Microarray Selection of Cooperative Peptides for Modulating Enzyme Activities. Microarrays. 2017 Jun;6(2):8.
- Stafford P, Halperin R, Legutki JB, Magee DM, Galgiani J, et al. Physical characterization of the ‘Immunosignaturing Effect’. Mol Cell Proteomics. 2012 Apr;11(4):M111.011593.
- Halperin RF, Stafford P, Legutki JB, Johnston SA. Exploring
antibody recognition of sequence space through randomsequence peptide microarrays. Mol Cell Proteomics. 2011
Mar;10(3):M110.000786.
- Geysen HM, Rodda SJ, Mason TJ, Tribbick G, Schoofs. Strategies
for epitope analysis using peptide synthesis. J Immunol Methods.
1987 Sep;102(2):259-274.
- Huang J, Ru B, Zhu P, Nie F, Yang J, et al. MimoDB 2.0: a mimotope
database and beyond. Nucleic Acids Res. 2012 Jan;40:D271-D277.
- Navalkar K, Magee DM, Galgiani J, Cichacz Z, Johnston SA, et al.
Application of immunosignatures to diagnosis of Valley Fever.
Clin Vaccine Immunol. 2014 Aug;21(8):1169-1177.
- Restrepo L, Stafford P, Magee DM, Johnston SA. Application of
immunosignatures to the assessment of Alzheimer’s disease. Ann
Neurol. 2011 Aug;70(2):286-295.
- Singh S, Stafford P, Schlauch KA, Tillett RR, Gollery M, et al. Humoral
Immunity Profiling of Subjects with Myalgic Encephalomyelitis
Using a Random Peptide Microarray Differentiates Cases from
Controls with High Specificity and Sensitivity. Mol Neurobiol.
2018 Jan;55(1):633-641.
- Stafford P, Cichacz Z, Woodbury NW, Johnston SA.
Immunosignature system for diagnosis of cancer. PNAS. 2014
Jul;111(30):E3072-E3080.
- Rodi DJ, Soares AS, Makowski L. Quantitative Assessment of
Peptide Sequence Diversity in M13 Combinatorial Peptide Phage
Display Libraries. J Mol Biol. 2002 Oct;322(5):1039-1052.
- Legutki JB, Zhao ZG, Greving M, Woodbury N, Johnston SA, et al.
Scalable high-density peptide arrays for comprehensive health
monitoring. Nat Commun. 2014 Sep;5:4785.
- Richer J, Johnston SA, Stafford P. Epitope Identification from
Fixed-complexity Random-sequence Peptide Microarrays. Mol
Cell Proteomics. 2015 Jan;14(1):136-147.
- Halperin R, Stafford P, Johnston SA. GuiTope: an application for
mapping random-sequence peptides to protein sequences. BMC
Bioinformatics. 2012;13:1.
- Donnell B, Maurer A, Papandreou-Suppappola A, Stafford P. TimeFrequency Analysis of Peptide Microarray Data: Application
to Brain Cancer Immunosignatures. Cancer Inform. 2015
Jun;14(Suppl 2):219-233.
- Zwick MB, Shen J, Scott JK. Phage-displayed peptide libraries.
Curr Opin Biotechnol. 1998 Aug;9(4):427-436.
- Restrepo L, Stafford P, Johnston SA. Feasibility of an early Alzheimer’s disease immunosignature diagnostic test. J Neuroimmunol. 2013 Jan;254(1-2):154-160.
- Williams S, Stafford P, Hoffman S. Diagnosis and early detection
of CNS-SLE in MRL/lpr mice using peptide microarrays. BMC
Immunol. 2014 Jun;15:23.
- Hughes A, Cichacz Z, Scheck AC, Coons SW, Johnston SA, et al.
Immunosignaturing can detect products from molecular markers
in brain cancer. PLoS ONE. 2012;7(7):e40201.
- Stafford P, Wrapp D, Johnston SA. General Assessment of
Humoral Activity in Healthy Humans. Mol Cell Proteomics. 2016
May;15(5):1610-1621.
- Wang L, Whittemore K, Johnston SA, Stafford P. Entropy is a Simple
Measure of the Antibody Profile and is an Indicator of Health
Status: A Proof of Concept. Scientific Reports. 2017 Dec;7:18060.
- Young JD, Huang AS, Ariel N, Bruins JB, Ng D, et al. Coupling
Efficiencies of Amino Acids in the Solid Phase Synthesis of
Peptides. Pept Res. 1990 Aug;3(4):194-200.
- Gao X, Pellois JP, Na Y, Kim Y, Gulari E, et al. High density peptide
microarrays. In situ synthesis and applications. Mol Divers.
2004;8(3):177-187.
- Gao X, Gulari E, Zhou X. In situ synthesis of oligonucleotide
microarrays. Biopolymers. 2004 Apr;73:579-596.
- Fodor S, Read J, Pirrung M, Stryer L, Lu A, et al. Light-directed,
spatially addressable parallel chemical synthesis. Science. 1991
Feb;251(4995):767-773.
- Choung RS, Marietta EV, Van Dyke CT, Brantner TL, Rajasekaran
J, et al. Determination of B-Cell Epitopes in Patients with
Celiac Disease: Peptide Microarrays. PLoS One. 2016
Jan;11(1):e0147777.
- Pellois JP, Zhou X, Srivannavit O, Zhou T, Gulari E, et al. Individually
addressable parallel peptide synthesis on microchips. Nat
Biotechnol. 2002 Sep;20(9):922-926.
- Beyer M, Nesterov A, Block I, König K, Felgenhauer T, et al.
Combinatorial synthesis of peptide arrays onto a microchip.
Science. 2007 Dec;318(5858):1888.
- Maurer K, McShea A, Strathmann M, Dill K. The removal of the t-BOC group by electrochemically generated acid and use of an addressable electrode array for peptide synthesis. J Comb Chem. 2005 Sep-Oct;7(5):637-640.
- Agarwal PB, Pawar S, Reddy SM, Mishra P, Agarwal A. Reusable
silicon shadow mask with sub-5 μm gap for low cost patterning.
Sensors and Actuators A: Physical. 2016 May;242:67-72.
- Kay BK, Adey NB, Yun-Sheng H, Manfredi JP, Mataragnon AH, et al.
An M13 phage library displaying random 38-amino-acid peptides
as a source of novel sequences with affinity to selected targets.
Gene. 1993 Jun;128(1):59-65.
- Noren KA, Noren CJ. Construction of High-Complexity
Combinatorial Phage Display Peptide Libraries. Methods.
2001;23(20:169-178.
- Cai C, Dai X, Zhu Y, Lian M, Xiao F, et al. A specific RAGE-binding peptide biopanning from phage display random peptide library that ameliorates symptoms in amyloid β peptidemediated neuronal disorder. Appl Microbiol Biotechnol. 2016 Jan;100(2):825-835.
- Ge M, Yan A, Luo W, Hu YF, Li R-C, et al. Epitope screening of the
PCV2 Cap protein by use of a random peptide-displayed library
and polyclonal antibody. Virus Res. 2013 Oct;177(1):103-107.
- Chin CF, Tan SJ, Gan CY, Lim TS. Identification of Peptide Based
Inhibitors for α-Amylase by Phage Display. Int J Pept Res Ther.
2015 Feb;21(3):237-242.
- Kukreja M, Johnston SA, Stafford P. Immunosignaturing
microarrays distinguish antibody profiles of related pancreatic
diseases. J Proteomics Bioinform. 2012 Jan;S6:001.
- Legutki JB, Johnston SA (2013) Immunosignatures can predict
vaccine efficacy. PNAS. 2013 Nov;110(46):18614-18619.
- Boltz KW, Nagaraj V, Svarovsky S, Lake D, Stafford P, et al. High
throughput technology for the identification and characterization
of glycan binding peptides. Glycobiology. 2006;16(3):1155-1156.
- Kuznetsov IB. Identification of non-random sequence
properties in groups of signature peptides obtained in random
sequence peptide microarray experiments. Biopolymers. 2016
May;106(3):318-29.
- Giancarlo R, Scaturro D, Utro F. Textual data compression
in computational biology: a synopsis. Bioinformatics.
2009;25(13):1575-1586.