US20030171274A1 - Antimicrobial proteins - Google Patents
Antimicrobial proteins Download PDFInfo
- Publication number
- US20030171274A1 US20030171274A1 US10/147,095 US14709502A US2003171274A1 US 20030171274 A1 US20030171274 A1 US 20030171274A1 US 14709502 A US14709502 A US 14709502A US 2003171274 A1 US2003171274 A1 US 2003171274A1
- Authority
- US
- United States
- Prior art keywords
- glu
- arg
- gln
- gly
- leu
- Prior art date
- Legal status (The legal status is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the status listed.)
- Abandoned
Links
- 108090000623 proteins and genes Proteins 0.000 title claims abstract description 267
- 102000004169 proteins and genes Human genes 0.000 title claims abstract description 236
- 230000000845 anti-microbial effect Effects 0.000 title claims abstract description 118
- 241000196324 Embryophyta Species 0.000 claims abstract description 106
- 239000000203 mixture Substances 0.000 claims abstract description 12
- 241001465754 Metazoa Species 0.000 claims abstract description 11
- 230000000813 microbial effect Effects 0.000 claims abstract description 9
- 206010061217 Infestation Diseases 0.000 claims abstract description 5
- 235000018102 proteins Nutrition 0.000 claims description 224
- 239000012634 fragment Substances 0.000 claims description 78
- XUJNEKJLAYXESH-REOHCLBHSA-N L-Cysteine Chemical compound SC[C@H](N)C(O)=O XUJNEKJLAYXESH-REOHCLBHSA-N 0.000 claims description 64
- 108090000765 processed proteins & peptides Proteins 0.000 claims description 59
- 238000000034 method Methods 0.000 claims description 45
- 108020004414 DNA Proteins 0.000 claims description 41
- 102000004196 processed proteins & peptides Human genes 0.000 claims description 33
- 150000001413 amino acids Chemical group 0.000 claims description 31
- 230000014509 gene expression Effects 0.000 claims description 31
- 235000018417 cysteine Nutrition 0.000 claims description 28
- 230000009261 transgenic effect Effects 0.000 claims description 26
- XUJNEKJLAYXESH-UHFFFAOYSA-N cysteine Natural products SCC(N)C(O)=O XUJNEKJLAYXESH-UHFFFAOYSA-N 0.000 claims description 24
- 125000000151 cysteine group Chemical group N[C@@H](CS)C(=O)* 0.000 claims description 23
- 229920001184 polypeptide Polymers 0.000 claims description 20
- 229920000742 Cotton Polymers 0.000 claims description 19
- 241000219146 Gossypium Species 0.000 claims description 19
- COLNVLDHVKWLRT-QMMMGPOBSA-N L-phenylalanine Chemical compound OC(=O)[C@@H](N)CC1=CC=CC=C1 COLNVLDHVKWLRT-QMMMGPOBSA-N 0.000 claims description 14
- OUYCCCASQSFEME-QMMMGPOBSA-N L-tyrosine Chemical compound OC(=O)[C@@H](N)CC1=CC=C(O)C=C1 OUYCCCASQSFEME-QMMMGPOBSA-N 0.000 claims description 14
- OUYCCCASQSFEME-UHFFFAOYSA-N tyrosine Natural products OC(=O)C(N)CC1=CC=C(O)C=C1 OUYCCCASQSFEME-UHFFFAOYSA-N 0.000 claims description 13
- 239000004599 antimicrobial Substances 0.000 claims description 11
- COLNVLDHVKWLRT-UHFFFAOYSA-N phenylalanine Natural products OC(=O)C(N)CC1=CC=CC=C1 COLNVLDHVKWLRT-UHFFFAOYSA-N 0.000 claims description 10
- 244000061176 Nicotiana tabacum Species 0.000 claims description 9
- 235000002637 Nicotiana tabacum Nutrition 0.000 claims description 9
- 240000008042 Zea mays Species 0.000 claims description 9
- 235000016383 Zea mays subsp huehuetenangensis Nutrition 0.000 claims description 9
- 235000002017 Zea mays subsp mays Nutrition 0.000 claims description 9
- 238000009826 distribution Methods 0.000 claims description 9
- 235000009973 maize Nutrition 0.000 claims description 9
- 238000009630 liquid culture Methods 0.000 claims description 8
- 244000068988 Glycine max Species 0.000 claims description 7
- 235000010469 Glycine max Nutrition 0.000 claims description 7
- 235000017060 Arachis glabrata Nutrition 0.000 claims description 6
- 244000105624 Arachis hypogaea Species 0.000 claims description 6
- 235000010777 Arachis hypogaea Nutrition 0.000 claims description 6
- 235000018262 Arachis monticola Nutrition 0.000 claims description 6
- 125000000539 amino acid group Chemical group 0.000 claims description 6
- 235000020232 peanut Nutrition 0.000 claims description 6
- 235000007340 Hordeum vulgare Nutrition 0.000 claims description 5
- 240000005979 Hordeum vulgare Species 0.000 claims description 5
- 239000000463 material Substances 0.000 claims description 5
- 239000003085 diluting agent Substances 0.000 claims description 4
- 239000000546 pharmaceutical excipient Substances 0.000 claims description 4
- 235000007319 Avena orientalis Nutrition 0.000 claims description 2
- 244000075850 Avena orientalis Species 0.000 claims description 2
- 235000014698 Brassica juncea var multisecta Nutrition 0.000 claims description 2
- 235000006008 Brassica napus var napus Nutrition 0.000 claims description 2
- 240000000385 Brassica napus var. napus Species 0.000 claims description 2
- 235000006618 Brassica rapa subsp oleifera Nutrition 0.000 claims description 2
- 235000004977 Brassica sinapistrum Nutrition 0.000 claims description 2
- 240000006497 Dianthus caryophyllus Species 0.000 claims description 2
- 235000009355 Dianthus caryophyllus Nutrition 0.000 claims description 2
- 235000003222 Helianthus annuus Nutrition 0.000 claims description 2
- 244000020551 Helianthus annuus Species 0.000 claims description 2
- 235000007688 Lycopersicon esculentum Nutrition 0.000 claims description 2
- 240000005561 Musa balbisiana Species 0.000 claims description 2
- 235000018290 Musa x paradisiaca Nutrition 0.000 claims description 2
- 240000004713 Pisum sativum Species 0.000 claims description 2
- 235000010582 Pisum sativum Nutrition 0.000 claims description 2
- 108020004511 Recombinant DNA Proteins 0.000 claims description 2
- 235000004789 Rosa xanthina Nutrition 0.000 claims description 2
- 241000109329 Rosa xanthina Species 0.000 claims description 2
- 240000003768 Solanum lycopersicum Species 0.000 claims description 2
- 244000061456 Solanum tuberosum Species 0.000 claims description 2
- 235000002595 Solanum tuberosum Nutrition 0.000 claims description 2
- 235000011684 Sorghum saccharatum Nutrition 0.000 claims description 2
- 235000021307 Triticum Nutrition 0.000 claims description 2
- 244000098338 Triticum aestivum Species 0.000 claims description 2
- 239000003937 drug carrier Substances 0.000 claims description 2
- 230000001850 reproductive effect Effects 0.000 claims description 2
- FWMNVWWHGCHHJJ-SKKKGAJSSA-N 4-amino-1-[(2r)-6-amino-2-[[(2r)-2-[[(2r)-2-[[(2r)-2-amino-3-phenylpropanoyl]amino]-3-phenylpropanoyl]amino]-4-methylpentanoyl]amino]hexanoyl]piperidine-4-carboxylic acid Chemical compound C([C@H](C(=O)N[C@H](CC(C)C)C(=O)N[C@H](CCCCN)C(=O)N1CCC(N)(CC1)C(O)=O)NC(=O)[C@H](N)CC=1C=CC=CC=1)C1=CC=CC=C1 FWMNVWWHGCHHJJ-SKKKGAJSSA-N 0.000 claims 1
- 240000006394 Sorghum bicolor Species 0.000 claims 1
- 240000007575 Macadamia integrifolia Species 0.000 abstract description 29
- 235000018330 Macadamia integrifolia Nutrition 0.000 abstract description 25
- 101000924450 Macadamia integrifolia Vicilin-like antimicrobial peptides 2-1 Proteins 0.000 description 52
- 101000924449 Macadamia integrifolia Vicilin-like antimicrobial peptides 2-2 Proteins 0.000 description 52
- 101000757796 Macadamia integrifolia Vicilin-like antimicrobial peptides 2-3 Proteins 0.000 description 52
- 244000299461 Theobroma cacao Species 0.000 description 31
- 239000013598 vector Substances 0.000 description 31
- 235000009470 Theobroma cacao Nutrition 0.000 description 30
- 229940024606 amino acid Drugs 0.000 description 30
- 235000001014 amino acid Nutrition 0.000 description 28
- 210000004027 cell Anatomy 0.000 description 27
- BUZMZDDKFCSKOT-CIUDSAMLSA-N Glu-Glu-Glu Chemical compound OC(=O)CC[C@H](N)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CCC(O)=O)C(O)=O BUZMZDDKFCSKOT-CIUDSAMLSA-N 0.000 description 26
- KOSRFJWDECSPRO-UHFFFAOYSA-N alpha-L-glutamyl-L-glutamic acid Natural products OC(=O)CCC(N)C(=O)NC(CCC(O)=O)C(O)=O KOSRFJWDECSPRO-UHFFFAOYSA-N 0.000 description 25
- WHUUTDBJXJRKMK-UHFFFAOYSA-N Glutamic acid Natural products OC(=O)C(N)CCC(O)=O WHUUTDBJXJRKMK-UHFFFAOYSA-N 0.000 description 22
- 101800000312 Antimicrobial peptide 2a Proteins 0.000 description 21
- DTQVDTLACAAQTR-UHFFFAOYSA-N Trifluoroacetic acid Chemical compound OC(=O)C(F)(F)F DTQVDTLACAAQTR-UHFFFAOYSA-N 0.000 description 20
- 125000003275 alpha amino acid group Chemical group 0.000 description 19
- 108010055341 glutamyl-glutamic acid Proteins 0.000 description 19
- 108091028043 Nucleic acid sequence Proteins 0.000 description 18
- PMGDADKJMCOXHX-UHFFFAOYSA-N L-Arginyl-L-glutamin-acetat Natural products NC(=N)NCCCC(N)C(=O)NC(CCC(N)=O)C(O)=O PMGDADKJMCOXHX-UHFFFAOYSA-N 0.000 description 17
- 108010068380 arginylarginine Proteins 0.000 description 17
- 238000000746 purification Methods 0.000 description 17
- WEVYAHXRMPXWCK-UHFFFAOYSA-N Acetonitrile Chemical compound CC#N WEVYAHXRMPXWCK-UHFFFAOYSA-N 0.000 description 16
- 108010016634 Seed Storage Proteins Proteins 0.000 description 16
- 101710196023 Vicilin Proteins 0.000 description 16
- 108010052670 arginyl-glutamyl-glutamic acid Proteins 0.000 description 15
- 230000000694 effects Effects 0.000 description 15
- YBYRMVIVWMBXKQ-UHFFFAOYSA-N phenylmethanesulfonyl fluoride Chemical compound FS(=O)(=O)CC1=CC=CC=C1 YBYRMVIVWMBXKQ-UHFFFAOYSA-N 0.000 description 15
- 108700042778 Antimicrobial Peptides Proteins 0.000 description 14
- 102000044503 Antimicrobial Peptides Human genes 0.000 description 14
- 239000000284 extract Substances 0.000 description 14
- 238000012360 testing method Methods 0.000 description 14
- 108010008355 arginyl-glutamine Proteins 0.000 description 13
- 241000208467 Macadamia Species 0.000 description 12
- XSQUKJJJFZCRTK-UHFFFAOYSA-N Urea Chemical compound NC(N)=O XSQUKJJJFZCRTK-UHFFFAOYSA-N 0.000 description 12
- 241000894007 species Species 0.000 description 12
- 241000233866 Fungi Species 0.000 description 11
- PKVWNYGXMNWJSI-CIUDSAMLSA-N Gln-Gln-Gln Chemical compound [H]N[C@@H](CCC(N)=O)C(=O)N[C@@H](CCC(N)=O)C(=O)N[C@@H](CCC(N)=O)C(O)=O PKVWNYGXMNWJSI-CIUDSAMLSA-N 0.000 description 11
- YBAFDPFAUTYYRW-UHFFFAOYSA-N N-L-alpha-glutamyl-L-leucine Natural products CC(C)CC(C(O)=O)NC(=O)C(N)CCC(O)=O YBAFDPFAUTYYRW-UHFFFAOYSA-N 0.000 description 11
- 239000000047 product Substances 0.000 description 11
- 108091035707 Consensus sequence Proteins 0.000 description 10
- XHWLNISLUFEWNS-CIUDSAMLSA-N Glu-Gln-Gln Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N[C@@H](CCC(N)=O)C(=O)N[C@@H](CCC(N)=O)C(O)=O XHWLNISLUFEWNS-CIUDSAMLSA-N 0.000 description 10
- RJIVPOXLQFJRTG-LURJTMIESA-N Gly-Arg-Gly Chemical compound OC(=O)CNC(=O)[C@@H](NC(=O)CN)CCCN=C(N)N RJIVPOXLQFJRTG-LURJTMIESA-N 0.000 description 10
- SITLTJHOQZFJGG-UHFFFAOYSA-N N-L-alpha-glutamyl-L-valine Natural products CC(C)C(C(O)=O)NC(=O)C(N)CCC(O)=O SITLTJHOQZFJGG-UHFFFAOYSA-N 0.000 description 10
- 108091034117 Oligonucleotide Proteins 0.000 description 10
- 240000005924 Stenocarpus sinuatus Species 0.000 description 10
- 108010013835 arginine glutamate Proteins 0.000 description 10
- 238000004166 bioassay Methods 0.000 description 10
- 239000002299 complementary DNA Substances 0.000 description 10
- 238000010828 elution Methods 0.000 description 10
- 239000002773 nucleotide Substances 0.000 description 10
- 125000003729 nucleotide group Chemical group 0.000 description 10
- 101800000311 Antimicrobial peptide 2b Proteins 0.000 description 9
- KJGNDQCYBNBXDA-GUBZILKMSA-N Arg-Arg-Cys Chemical compound C(C[C@@H](C(=O)N[C@@H](CCCN=C(N)N)C(=O)N[C@@H](CS)C(=O)O)N)CN=C(N)N KJGNDQCYBNBXDA-GUBZILKMSA-N 0.000 description 9
- 241000588724 Escherichia coli Species 0.000 description 9
- 108010076504 Protein Sorting Signals Proteins 0.000 description 9
- 238000005341 cation exchange Methods 0.000 description 9
- 238000003776 cleavage reaction Methods 0.000 description 9
- 108010009298 lysylglutamic acid Proteins 0.000 description 9
- 230000002441 reversible effect Effects 0.000 description 9
- 238000002415 sodium dodecyl sulfate polyacrylamide gel electrophoresis Methods 0.000 description 9
- KBBKCNHWCDJPGN-GUBZILKMSA-N Arg-Gln-Gln Chemical compound [H]N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CCC(N)=O)C(=O)N[C@@H](CCC(N)=O)C(O)=O KBBKCNHWCDJPGN-GUBZILKMSA-N 0.000 description 8
- CYXCAHZVPFREJD-LURJTMIESA-N Arg-Gly-Gly Chemical compound NC(=N)NCCC[C@H](N)C(=O)NCC(=O)NCC(O)=O CYXCAHZVPFREJD-LURJTMIESA-N 0.000 description 8
- 108091026890 Coding region Proteins 0.000 description 8
- KCXVZYZYPLLWCC-UHFFFAOYSA-N EDTA Chemical compound OC(=O)CN(CC(O)=O)CCN(CC(O)=O)CC(O)=O KCXVZYZYPLLWCC-UHFFFAOYSA-N 0.000 description 8
- DHDOADIPGZTAHT-YUMQZZPRSA-N Gly-Glu-Arg Chemical compound NCC(=O)N[C@@H](CCC(O)=O)C(=O)N[C@H](C(O)=O)CCCN=C(N)N DHDOADIPGZTAHT-YUMQZZPRSA-N 0.000 description 8
- 241000880493 Leptailurus serval Species 0.000 description 8
- 101710093543 Probable non-specific lipid-transfer protein Proteins 0.000 description 8
- 230000001413 cellular effect Effects 0.000 description 8
- 108010049041 glutamylalanine Proteins 0.000 description 8
- 239000013612 plasmid Substances 0.000 description 8
- 230000007017 scission Effects 0.000 description 8
- 101800000314 Antimicrobial peptide 2d Proteins 0.000 description 7
- UXJCMQFPDWCHKX-DCAQKATOSA-N Arg-Arg-Glu Chemical compound NC(N)=NCCC[C@H](N)C(=O)N[C@@H](CCCN=C(N)N)C(=O)N[C@@H](CCC(O)=O)C(O)=O UXJCMQFPDWCHKX-DCAQKATOSA-N 0.000 description 7
- KWUSGAIFNHQCBY-DCAQKATOSA-N Gln-Arg-Arg Chemical compound NC(=O)CC[C@H](N)C(=O)N[C@@H](CCCN=C(N)N)C(=O)N[C@@H](CCCN=C(N)N)C(O)=O KWUSGAIFNHQCBY-DCAQKATOSA-N 0.000 description 7
- ILGFBUGLBSAQQB-GUBZILKMSA-N Glu-Glu-Arg Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CCCNC(N)=N)C(O)=O ILGFBUGLBSAQQB-GUBZILKMSA-N 0.000 description 7
- SJPMNHCEWPTRBR-BQBZGAKWSA-N Glu-Glu-Gly Chemical compound OC(=O)CC[C@H](N)C(=O)N[C@@H](CCC(O)=O)C(=O)NCC(O)=O SJPMNHCEWPTRBR-BQBZGAKWSA-N 0.000 description 7
- KDXKERNSBIXSRK-UHFFFAOYSA-N Lysine Natural products NCCCCC(N)C(O)=O KDXKERNSBIXSRK-UHFFFAOYSA-N 0.000 description 7
- 108010002311 N-glycylglutamic acid Proteins 0.000 description 7
- FISHYTLIMUYTQY-GUBZILKMSA-N Pro-Gln-Gln Chemical compound NC(=O)CC[C@@H](C(O)=O)NC(=O)[C@H](CCC(N)=O)NC(=O)[C@@H]1CCCN1 FISHYTLIMUYTQY-GUBZILKMSA-N 0.000 description 7
- 238000002835 absorbance Methods 0.000 description 7
- 230000000843 anti-fungal effect Effects 0.000 description 7
- 108010093581 aspartyl-proline Proteins 0.000 description 7
- 108010038633 aspartylglutamate Proteins 0.000 description 7
- 230000015572 biosynthetic process Effects 0.000 description 7
- 238000006243 chemical reaction Methods 0.000 description 7
- 238000010367 cloning Methods 0.000 description 7
- 108010063718 gamma-glutamylaspartic acid Proteins 0.000 description 7
- 108010078144 glutaminyl-glycine Proteins 0.000 description 7
- 108010010147 glycylglutamine Proteins 0.000 description 7
- 230000005764 inhibitory process Effects 0.000 description 7
- BPHPUYQFMNQIOC-NXRLNHOXSA-N isopropyl beta-D-thiogalactopyranoside Chemical compound CC(C)S[C@@H]1O[C@H](CO)[C@H](O)[C@H](O)[C@H]1O BPHPUYQFMNQIOC-NXRLNHOXSA-N 0.000 description 7
- 230000036961 partial effect Effects 0.000 description 7
- 238000012216 screening Methods 0.000 description 7
- 238000001262 western blot Methods 0.000 description 7
- AWZKCUCQJNTBAD-SRVKXCTJSA-N Ala-Leu-Lys Chemical compound C[C@H](N)C(=O)N[C@@H](CC(C)C)C(=O)N[C@H](C(O)=O)CCCCN AWZKCUCQJNTBAD-SRVKXCTJSA-N 0.000 description 6
- IASNWHAGGYTEKX-IUCAKERBSA-N Arg-Arg-Gly Chemical compound NC(N)=NCCC[C@H](N)C(=O)N[C@@H](CCCN=C(N)N)C(=O)NCC(O)=O IASNWHAGGYTEKX-IUCAKERBSA-N 0.000 description 6
- HQIZDMIGUJOSNI-IUCAKERBSA-N Arg-Gly-Arg Chemical compound N[C@@H](CCCNC(N)=N)C(=O)NCC(=O)N[C@@H](CCCNC(N)=N)C(O)=O HQIZDMIGUJOSNI-IUCAKERBSA-N 0.000 description 6
- HNVFSTLPVJWIDV-CIUDSAMLSA-N Glu-Glu-Gln Chemical compound OC(=O)CC[C@H](N)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CCC(N)=O)C(O)=O HNVFSTLPVJWIDV-CIUDSAMLSA-N 0.000 description 6
- DHMQDGOQFOQNFH-UHFFFAOYSA-N Glycine Chemical compound NCC(O)=O DHMQDGOQFOQNFH-UHFFFAOYSA-N 0.000 description 6
- KZNQNBZMBZJQJO-UHFFFAOYSA-N N-glycyl-L-proline Natural products NCC(=O)N1CCCC1C(O)=O KZNQNBZMBZJQJO-UHFFFAOYSA-N 0.000 description 6
- 108010088535 Pep-1 peptide Proteins 0.000 description 6
- CEXFELBFVHLYDZ-XGEHTFHBSA-N Thr-Arg-Ser Chemical compound [H]N[C@@H]([C@@H](C)O)C(=O)N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CO)C(O)=O CEXFELBFVHLYDZ-XGEHTFHBSA-N 0.000 description 6
- WVRUKYLYMFGKAN-IHRRRGAJSA-N Tyr-Glu-Glu Chemical compound OC(=O)CC[C@@H](C(O)=O)NC(=O)[C@H](CCC(O)=O)NC(=O)[C@@H](N)CC1=CC=C(O)C=C1 WVRUKYLYMFGKAN-IHRRRGAJSA-N 0.000 description 6
- 230000002378 acidificating effect Effects 0.000 description 6
- 108010024078 alanyl-glycyl-serine Proteins 0.000 description 6
- 108010047495 alanylglycine Proteins 0.000 description 6
- 108010009111 arginyl-glycyl-glutamic acid Proteins 0.000 description 6
- 108010062796 arginyllysine Proteins 0.000 description 6
- 125000003118 aryl group Chemical group 0.000 description 6
- 108010040443 aspartyl-aspartic acid Proteins 0.000 description 6
- 239000004202 carbamide Substances 0.000 description 6
- 201000010099 disease Diseases 0.000 description 6
- 208000037265 diseases, disorders, signs and symptoms Diseases 0.000 description 6
- 230000002538 fungal effect Effects 0.000 description 6
- 239000000499 gel Substances 0.000 description 6
- 230000003993 interaction Effects 0.000 description 6
- 244000052769 pathogen Species 0.000 description 6
- 238000012545 processing Methods 0.000 description 6
- 108010004914 prolylarginine Proteins 0.000 description 6
- 108010053725 prolylvaline Proteins 0.000 description 6
- 238000004007 reversed phase HPLC Methods 0.000 description 6
- 238000000926 separation method Methods 0.000 description 6
- 238000012163 sequencing technique Methods 0.000 description 6
- 230000006641 stabilisation Effects 0.000 description 6
- 108091032973 (ribonucleotides)n+m Proteins 0.000 description 5
- 241000589155 Agrobacterium tumefaciens Species 0.000 description 5
- RYRQZJVFDVWURI-SRVKXCTJSA-N Arg-Gln-His Chemical compound C1=C(NC=N1)C[C@@H](C(=O)O)NC(=O)[C@H](CCC(=O)N)NC(=O)[C@H](CCCN=C(N)N)N RYRQZJVFDVWURI-SRVKXCTJSA-N 0.000 description 5
- PNQWAUXQDBIJDY-GUBZILKMSA-N Arg-Glu-Glu Chemical compound [H]N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CCC(O)=O)C(O)=O PNQWAUXQDBIJDY-GUBZILKMSA-N 0.000 description 5
- UFBURHXMKFQVLM-CIUDSAMLSA-N Arg-Glu-Ser Chemical compound [H]N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CO)C(O)=O UFBURHXMKFQVLM-CIUDSAMLSA-N 0.000 description 5
- AUFHLLPVPSMEOG-YUMQZZPRSA-N Arg-Gly-Glu Chemical compound NC(N)=NCCC[C@H](N)C(=O)NCC(=O)N[C@@H](CCC(O)=O)C(O)=O AUFHLLPVPSMEOG-YUMQZZPRSA-N 0.000 description 5
- MSILNNHVVMMTHZ-UWVGGRQHSA-N Arg-His-Gly Chemical compound NC(N)=NCCC[C@H](N)C(=O)N[C@H](C(=O)NCC(O)=O)CC1=CN=CN1 MSILNNHVVMMTHZ-UWVGGRQHSA-N 0.000 description 5
- CXFUMJQFZVCETK-FXQIFTODSA-N Gln-Cys-Gln Chemical compound NC(=O)CC[C@H](N)C(=O)N[C@@H](CS)C(=O)N[C@@H](CCC(N)=O)C(O)=O CXFUMJQFZVCETK-FXQIFTODSA-N 0.000 description 5
- NKCZYEDZTKOFBG-GUBZILKMSA-N Gln-Gln-Arg Chemical compound [H]N[C@@H](CCC(N)=O)C(=O)N[C@@H](CCC(N)=O)C(=O)N[C@@H](CCCNC(N)=N)C(O)=O NKCZYEDZTKOFBG-GUBZILKMSA-N 0.000 description 5
- ZQPOVSJFBBETHQ-CIUDSAMLSA-N Gln-Glu-Gln Chemical compound [H]N[C@@H](CCC(N)=O)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CCC(N)=O)C(O)=O ZQPOVSJFBBETHQ-CIUDSAMLSA-N 0.000 description 5
- 108010044091 Globulins Proteins 0.000 description 5
- 102000006395 Globulins Human genes 0.000 description 5
- XHUCVVHRLNPZSZ-CIUDSAMLSA-N Glu-Gln-Glu Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N[C@@H](CCC(N)=O)C(=O)N[C@@H](CCC(O)=O)C(O)=O XHUCVVHRLNPZSZ-CIUDSAMLSA-N 0.000 description 5
- SJJHXJDSNQJMMW-SRVKXCTJSA-N Glu-Lys-Arg Chemical compound OC(=O)CC[C@H](N)C(=O)N[C@@H](CCCCN)C(=O)N[C@@H](CCCN=C(N)N)C(O)=O SJJHXJDSNQJMMW-SRVKXCTJSA-N 0.000 description 5
- TWYFJOHWGCCRIR-DCAQKATOSA-N Glu-Pro-Arg Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N1CCC[C@H]1C(=O)N[C@@H](CCCNC(N)=N)C(O)=O TWYFJOHWGCCRIR-DCAQKATOSA-N 0.000 description 5
- HAPWZEVRQYGLSG-IUCAKERBSA-N His-Gly-Glu Chemical compound [H]N[C@@H](CC1=CNC=N1)C(=O)NCC(=O)N[C@@H](CCC(O)=O)C(O)=O HAPWZEVRQYGLSG-IUCAKERBSA-N 0.000 description 5
- TZCGZYWNIDZZMR-UHFFFAOYSA-N Ile-Arg-Ala Natural products CCC(C)C(N)C(=O)NC(C(=O)NC(C)C(O)=O)CCCN=C(N)N TZCGZYWNIDZZMR-UHFFFAOYSA-N 0.000 description 5
- HUORUFRRJHELPD-MNXVOIDGSA-N Ile-Leu-Glu Chemical compound CC[C@H](C)[C@@H](C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](CCC(=O)O)C(=O)O)N HUORUFRRJHELPD-MNXVOIDGSA-N 0.000 description 5
- QNBVTHNJGCOVFA-AVGNSLFASA-N Leu-Leu-Glu Chemical compound CC(C)C[C@H](N)C(=O)N[C@@H](CC(C)C)C(=O)N[C@H](C(O)=O)CCC(O)=O QNBVTHNJGCOVFA-AVGNSLFASA-N 0.000 description 5
- XMBSYZWANAQXEV-UHFFFAOYSA-N N-alpha-L-glutamyl-L-phenylalanine Natural products OC(=O)CCC(N)C(=O)NC(C(O)=O)CC1=CC=CC=C1 XMBSYZWANAQXEV-UHFFFAOYSA-N 0.000 description 5
- 238000012408 PCR amplification Methods 0.000 description 5
- WKLMCMXFMQEKCX-SLFFLAALSA-N Phe-Phe-Pro Chemical compound C1C[C@@H](N(C1)C(=O)[C@H](CC2=CC=CC=C2)NC(=O)[C@H](CC3=CC=CC=C3)N)C(=O)O WKLMCMXFMQEKCX-SLFFLAALSA-N 0.000 description 5
- KIZQGKLMXKGDIV-BQBZGAKWSA-N Pro-Ala-Gly Chemical compound OC(=O)CNC(=O)[C@H](C)NC(=O)[C@@H]1CCCN1 KIZQGKLMXKGDIV-BQBZGAKWSA-N 0.000 description 5
- AEGUWTFAQQWVLC-BQBZGAKWSA-N Ser-Gly-Arg Chemical compound [H]N[C@@H](CO)C(=O)NCC(=O)N[C@@H](CCCNC(N)=N)C(O)=O AEGUWTFAQQWVLC-BQBZGAKWSA-N 0.000 description 5
- AYHSJESDFKREAR-KKUMJFAQSA-N Tyr-Asn-Leu Chemical compound CC(C)C[C@@H](C(O)=O)NC(=O)[C@H](CC(N)=O)NC(=O)[C@@H](N)CC1=CC=C(O)C=C1 AYHSJESDFKREAR-KKUMJFAQSA-N 0.000 description 5
- AZSHAZJLOZQYAY-FXQIFTODSA-N Val-Ala-Ser Chemical compound CC(C)[C@H](N)C(=O)N[C@@H](C)C(=O)N[C@@H](CO)C(O)=O AZSHAZJLOZQYAY-FXQIFTODSA-N 0.000 description 5
- 108010044940 alanylglutamine Proteins 0.000 description 5
- 238000004458 analytical method Methods 0.000 description 5
- 238000013459 approach Methods 0.000 description 5
- 108010001271 arginyl-glutamyl-arginine Proteins 0.000 description 5
- 230000003115 biocidal effect Effects 0.000 description 5
- 230000008859 change Effects 0.000 description 5
- 238000001514 detection method Methods 0.000 description 5
- 230000001747 exhibiting effect Effects 0.000 description 5
- 108010050848 glycylleucine Proteins 0.000 description 5
- 108010045383 histidyl-glycyl-glutamic acid Proteins 0.000 description 5
- 108010036413 histidylglycine Proteins 0.000 description 5
- 244000000010 microbial pathogen Species 0.000 description 5
- 108010073025 phenylalanylphenylalanine Proteins 0.000 description 5
- 108091008146 restriction endonucleases Proteins 0.000 description 5
- 125000002653 sulfanylmethyl group Chemical group [H]SC([H])([H])[*] 0.000 description 5
- HZAXFHJVJLSVMW-UHFFFAOYSA-N 2-Aminoethan-1-ol Chemical compound NCCO HZAXFHJVJLSVMW-UHFFFAOYSA-N 0.000 description 4
- CJQAEJMHBAOQHA-DLOVCJGASA-N Ala-Phe-Asn Chemical compound C[C@@H](C(=O)N[C@@H](CC1=CC=CC=C1)C(=O)N[C@@H](CC(=O)N)C(=O)O)N CJQAEJMHBAOQHA-DLOVCJGASA-N 0.000 description 4
- QIWYWCYNUMJBTC-CIUDSAMLSA-N Arg-Cys-Gln Chemical compound [H]N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CS)C(=O)N[C@@H](CCC(N)=O)C(O)=O QIWYWCYNUMJBTC-CIUDSAMLSA-N 0.000 description 4
- GFMWTFHOZGLTLC-AVGNSLFASA-N Arg-His-Met Chemical compound [H]N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CC1=CNC=N1)C(=O)N[C@@H](CCSC)C(O)=O GFMWTFHOZGLTLC-AVGNSLFASA-N 0.000 description 4
- PJOPLXOCKACMLK-KKUMJFAQSA-N Arg-Tyr-Glu Chemical compound [H]N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CC1=CC=C(O)C=C1)C(=O)N[C@@H](CCC(O)=O)C(O)=O PJOPLXOCKACMLK-KKUMJFAQSA-N 0.000 description 4
- HJZLUGQGJWXJCJ-CIUDSAMLSA-N Asp-Pro-Gln Chemical compound [H]N[C@@H](CC(O)=O)C(=O)N1CCC[C@H]1C(=O)N[C@@H](CCC(N)=O)C(O)=O HJZLUGQGJWXJCJ-CIUDSAMLSA-N 0.000 description 4
- IJGRMHOSHXDMSA-UHFFFAOYSA-N Atomic nitrogen Chemical compound N#N IJGRMHOSHXDMSA-UHFFFAOYSA-N 0.000 description 4
- 241000894006 Bacteria Species 0.000 description 4
- BPHKULHWEIUDOB-FXQIFTODSA-N Cys-Gln-Gln Chemical compound SC[C@H](N)C(=O)N[C@@H](CCC(N)=O)C(=O)N[C@@H](CCC(N)=O)C(O)=O BPHKULHWEIUDOB-FXQIFTODSA-N 0.000 description 4
- LZRMPXRYLLTAJX-GUBZILKMSA-N Gln-Arg-Glu Chemical compound [H]N[C@@H](CCC(N)=O)C(=O)N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CCC(O)=O)C(O)=O LZRMPXRYLLTAJX-GUBZILKMSA-N 0.000 description 4
- SNLOOPZHAQDMJG-CIUDSAMLSA-N Gln-Glu-Glu Chemical compound NC(=O)CC[C@H](N)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CCC(O)=O)C(O)=O SNLOOPZHAQDMJG-CIUDSAMLSA-N 0.000 description 4
- GURIQZQSTBBHRV-SRVKXCTJSA-N Gln-Lys-Arg Chemical compound [H]N[C@@H](CCC(N)=O)C(=O)N[C@@H](CCCCN)C(=O)N[C@@H](CCCNC(N)=N)C(O)=O GURIQZQSTBBHRV-SRVKXCTJSA-N 0.000 description 4
- WIMVKDYAKRAUCG-IHRRRGAJSA-N Gln-Tyr-Glu Chemical compound C1=CC(=CC=C1C[C@@H](C(=O)N[C@@H](CCC(=O)O)C(=O)O)NC(=O)[C@H](CCC(=O)N)N)O WIMVKDYAKRAUCG-IHRRRGAJSA-N 0.000 description 4
- CGYDXNKRIMJMLV-GUBZILKMSA-N Glu-Arg-Glu Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CCC(O)=O)C(O)=O CGYDXNKRIMJMLV-GUBZILKMSA-N 0.000 description 4
- KKCUFHUTMKQQCF-SRVKXCTJSA-N Glu-Arg-Leu Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CC(C)C)C(O)=O KKCUFHUTMKQQCF-SRVKXCTJSA-N 0.000 description 4
- LGYZYFFDELZWRS-DCAQKATOSA-N Glu-Glu-Lys Chemical compound NCCCC[C@@H](C(O)=O)NC(=O)[C@H](CCC(O)=O)NC(=O)[C@@H](N)CCC(O)=O LGYZYFFDELZWRS-DCAQKATOSA-N 0.000 description 4
- 241000238631 Hexapoda Species 0.000 description 4
- HGCNKOLVKRAVHD-UHFFFAOYSA-N L-Met-L-Phe Natural products CSCCC(N)C(=O)NC(C(O)=O)CC1=CC=CC=C1 HGCNKOLVKRAVHD-UHFFFAOYSA-N 0.000 description 4
- FADYJNXDPBKVCA-UHFFFAOYSA-N L-Phenylalanyl-L-lysin Natural products NCCCCC(C(O)=O)NC(=O)C(N)CC1=CC=CC=C1 FADYJNXDPBKVCA-UHFFFAOYSA-N 0.000 description 4
- WNGVUZWBXZKQES-YUMQZZPRSA-N Leu-Ala-Gly Chemical compound CC(C)C[C@H](N)C(=O)N[C@@H](C)C(=O)NCC(O)=O WNGVUZWBXZKQES-YUMQZZPRSA-N 0.000 description 4
- KAFOIVJDVSZUMD-UHFFFAOYSA-N Leu-Gln-Gln Natural products CC(C)CC(N)C(=O)NC(CCC(N)=O)C(=O)NC(CCC(N)=O)C(O)=O KAFOIVJDVSZUMD-UHFFFAOYSA-N 0.000 description 4
- DSFYPIUSAMSERP-IHRRRGAJSA-N Leu-Leu-Arg Chemical compound CC(C)C[C@H](N)C(=O)N[C@@H](CC(C)C)C(=O)N[C@H](C(O)=O)CCCN=C(N)N DSFYPIUSAMSERP-IHRRRGAJSA-N 0.000 description 4
- PPGBXYKMUMHFBF-KATARQTJSA-N Leu-Ser-Thr Chemical compound [H]N[C@@H](CC(C)C)C(=O)N[C@@H](CO)C(=O)N[C@@H]([C@@H](C)O)C(O)=O PPGBXYKMUMHFBF-KATARQTJSA-N 0.000 description 4
- 235000003800 Macadamia tetraphylla Nutrition 0.000 description 4
- 241000221696 Sclerotinia sclerotiorum Species 0.000 description 4
- LWMQRHDTXHQQOV-MXAVVETBSA-N Ser-Ile-Phe Chemical compound [H]N[C@@H](CO)C(=O)N[C@@H]([C@@H](C)CC)C(=O)N[C@@H](CC1=CC=CC=C1)C(O)=O LWMQRHDTXHQQOV-MXAVVETBSA-N 0.000 description 4
- GVMUJUPXFQFBBZ-GUBZILKMSA-N Ser-Lys-Glu Chemical compound [H]N[C@@H](CO)C(=O)N[C@@H](CCCCN)C(=O)N[C@@H](CCC(O)=O)C(O)=O GVMUJUPXFQFBBZ-GUBZILKMSA-N 0.000 description 4
- FAPWRFPIFSIZLT-UHFFFAOYSA-M Sodium chloride Chemical compound [Na+].[Cl-] FAPWRFPIFSIZLT-UHFFFAOYSA-M 0.000 description 4
- CXWJFWAZIVWBOS-XQQFMLRXSA-N Val-Lys-Pro Chemical compound CC(C)[C@@H](C(=O)N[C@@H](CCCCN)C(=O)N1CCC[C@@H]1C(=O)O)N CXWJFWAZIVWBOS-XQQFMLRXSA-N 0.000 description 4
- JLCPHMBAVCMARE-UHFFFAOYSA-N [3-[[3-[[3-[[3-[[3-[[3-[[3-[[3-[[3-[[3-[[3-[[5-(2-amino-6-oxo-1H-purin-9-yl)-3-[[3-[[3-[[3-[[3-[[3-[[5-(2-amino-6-oxo-1H-purin-9-yl)-3-[[5-(2-amino-6-oxo-1H-purin-9-yl)-3-hydroxyoxolan-2-yl]methoxy-hydroxyphosphoryl]oxyoxolan-2-yl]methoxy-hydroxyphosphoryl]oxy-5-(5-methyl-2,4-dioxopyrimidin-1-yl)oxolan-2-yl]methoxy-hydroxyphosphoryl]oxy-5-(6-aminopurin-9-yl)oxolan-2-yl]methoxy-hydroxyphosphoryl]oxy-5-(6-aminopurin-9-yl)oxolan-2-yl]methoxy-hydroxyphosphoryl]oxy-5-(6-aminopurin-9-yl)oxolan-2-yl]methoxy-hydroxyphosphoryl]oxy-5-(6-aminopurin-9-yl)oxolan-2-yl]methoxy-hydroxyphosphoryl]oxyoxolan-2-yl]methoxy-hydroxyphosphoryl]oxy-5-(5-methyl-2,4-dioxopyrimidin-1-yl)oxolan-2-yl]methoxy-hydroxyphosphoryl]oxy-5-(4-amino-2-oxopyrimidin-1-yl)oxolan-2-yl]methoxy-hydroxyphosphoryl]oxy-5-(5-methyl-2,4-dioxopyrimidin-1-yl)oxolan-2-yl]methoxy-hydroxyphosphoryl]oxy-5-(5-methyl-2,4-dioxopyrimidin-1-yl)oxolan-2-yl]methoxy-hydroxyphosphoryl]oxy-5-(6-aminopurin-9-yl)oxolan-2-yl]methoxy-hydroxyphosphoryl]oxy-5-(6-aminopurin-9-yl)oxolan-2-yl]methoxy-hydroxyphosphoryl]oxy-5-(4-amino-2-oxopyrimidin-1-yl)oxolan-2-yl]methoxy-hydroxyphosphoryl]oxy-5-(4-amino-2-oxopyrimidin-1-yl)oxolan-2-yl]methoxy-hydroxyphosphoryl]oxy-5-(4-amino-2-oxopyrimidin-1-yl)oxolan-2-yl]methoxy-hydroxyphosphoryl]oxy-5-(6-aminopurin-9-yl)oxolan-2-yl]methoxy-hydroxyphosphoryl]oxy-5-(4-amino-2-oxopyrimidin-1-yl)oxolan-2-yl]methyl [5-(6-aminopurin-9-yl)-2-(hydroxymethyl)oxolan-3-yl] hydrogen phosphate Polymers Cc1cn(C2CC(OP(O)(=O)OCC3OC(CC3OP(O)(=O)OCC3OC(CC3O)n3cnc4c3nc(N)[nH]c4=O)n3cnc4c3nc(N)[nH]c4=O)C(COP(O)(=O)OC3CC(OC3COP(O)(=O)OC3CC(OC3COP(O)(=O)OC3CC(OC3COP(O)(=O)OC3CC(OC3COP(O)(=O)OC3CC(OC3COP(O)(=O)OC3CC(OC3COP(O)(=O)OC3CC(OC3COP(O)(=O)OC3CC(OC3COP(O)(=O)OC3CC(OC3COP(O)(=O)OC3CC(OC3COP(O)(=O)OC3CC(OC3COP(O)(=O)OC3CC(OC3COP(O)(=O)OC3CC(OC3COP(O)(=O)OC3CC(OC3COP(O)(=O)OC3CC(OC3COP(O)(=O)OC3CC(OC3COP(O)(=O)OC3CC(OC3CO)n3cnc4c(N)ncnc34)n3ccc(N)nc3=O)n3cnc4c(N)ncnc34)n3ccc(N)nc3=O)n3ccc(N)nc3=O)n3ccc(N)nc3=O)n3cnc4c(N)ncnc34)n3cnc4c(N)ncnc34)n3cc(C)c(=O)[nH]c3=O)n3cc(C)c(=O)[nH]c3=O)n3ccc(N)nc3=O)n3cc(C)c(=O)[nH]c3=O)n3cnc4c3nc(N)[nH]c4=O)n3cnc4c(N)ncnc34)n3cnc4c(N)ncnc34)n3cnc4c(N)ncnc34)n3cnc4c(N)ncnc34)O2)c(=O)[nH]c1=O JLCPHMBAVCMARE-UHFFFAOYSA-N 0.000 description 4
- 108010005233 alanylglutamic acid Proteins 0.000 description 4
- 108010091092 arginyl-glycyl-proline Proteins 0.000 description 4
- 108010060035 arginylproline Proteins 0.000 description 4
- 108010077245 asparaginyl-proline Proteins 0.000 description 4
- 238000005119 centrifugation Methods 0.000 description 4
- 239000000287 crude extract Substances 0.000 description 4
- 150000001945 cysteines Chemical class 0.000 description 4
- 238000005194 fractionation Methods 0.000 description 4
- 230000006870 function Effects 0.000 description 4
- 230000004927 fusion Effects 0.000 description 4
- 108010020688 glycylhistidine Proteins 0.000 description 4
- 108010077515 glycylproline Proteins 0.000 description 4
- 108010028295 histidylhistidine Proteins 0.000 description 4
- 108010025306 histidylleucine Proteins 0.000 description 4
- 230000002209 hydrophobic effect Effects 0.000 description 4
- 230000002401 inhibitory effect Effects 0.000 description 4
- 108010044056 leucyl-phenylalanine Proteins 0.000 description 4
- 239000012528 membrane Substances 0.000 description 4
- 108010051242 phenylalanylserine Proteins 0.000 description 4
- 108010029020 prolylglycine Proteins 0.000 description 4
- 229920005989 resin Polymers 0.000 description 4
- 239000011347 resin Substances 0.000 description 4
- 230000028327 secretion Effects 0.000 description 4
- 238000011105 stabilization Methods 0.000 description 4
- 239000006228 supernatant Substances 0.000 description 4
- 108010061238 threonyl-glycine Proteins 0.000 description 4
- 210000001519 tissue Anatomy 0.000 description 4
- 230000009466 transformation Effects 0.000 description 4
- 238000013519 translation Methods 0.000 description 4
- QTBSBXVTEAMEQO-UHFFFAOYSA-N Acetic acid Chemical compound CC(O)=O QTBSBXVTEAMEQO-UHFFFAOYSA-N 0.000 description 3
- 241000589158 Agrobacterium Species 0.000 description 3
- XCVRVWZTXPCYJT-BIIVOSGPSA-N Ala-Asn-Pro Chemical compound C[C@@H](C(=O)N[C@@H](CC(=O)N)C(=O)N1CCC[C@@H]1C(=O)O)N XCVRVWZTXPCYJT-BIIVOSGPSA-N 0.000 description 3
- VGPWRRFOPXVGOH-BYPYZUCNSA-N Ala-Gly-Gly Chemical compound C[C@H](N)C(=O)NCC(=O)NCC(O)=O VGPWRRFOPXVGOH-BYPYZUCNSA-N 0.000 description 3
- NBTGEURICRTMGL-WHFBIAKZSA-N Ala-Gly-Ser Chemical compound C[C@H](N)C(=O)NCC(=O)N[C@@H](CO)C(O)=O NBTGEURICRTMGL-WHFBIAKZSA-N 0.000 description 3
- LNNSWWRRYJLGNI-NAKRPEOUSA-N Ala-Ile-Val Chemical compound C[C@H](N)C(=O)N[C@@H]([C@@H](C)CC)C(=O)N[C@@H](C(C)C)C(O)=O LNNSWWRRYJLGNI-NAKRPEOUSA-N 0.000 description 3
- HHRAXZAYZFFRAM-CIUDSAMLSA-N Ala-Leu-Asn Chemical compound [H]N[C@@H](C)C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](CC(N)=O)C(O)=O HHRAXZAYZFFRAM-CIUDSAMLSA-N 0.000 description 3
- DHBKYZYFEXXUAK-ONGXEEELSA-N Ala-Phe-Gly Chemical compound OC(=O)CNC(=O)[C@@H](NC(=O)[C@@H](N)C)CC1=CC=CC=C1 DHBKYZYFEXXUAK-ONGXEEELSA-N 0.000 description 3
- VKKYFICVTYKFIO-CIUDSAMLSA-N Arg-Ala-Glu Chemical compound OC(=O)CC[C@@H](C(O)=O)NC(=O)[C@H](C)NC(=O)[C@@H](N)CCCN=C(N)N VKKYFICVTYKFIO-CIUDSAMLSA-N 0.000 description 3
- XPSGESXVBSQZPL-SRVKXCTJSA-N Arg-Arg-Arg Chemical compound NC(N)=NCCC[C@H](N)C(=O)N[C@@H](CCCN=C(N)N)C(=O)N[C@@H](CCCN=C(N)N)C(O)=O XPSGESXVBSQZPL-SRVKXCTJSA-N 0.000 description 3
- JGDGLDNAQJJGJI-AVGNSLFASA-N Arg-Arg-His Chemical compound C1=C(NC=N1)C[C@@H](C(=O)O)NC(=O)[C@H](CCCN=C(N)N)NC(=O)[C@H](CCCN=C(N)N)N JGDGLDNAQJJGJI-AVGNSLFASA-N 0.000 description 3
- UISQLSIBJKEJSS-GUBZILKMSA-N Arg-Arg-Ser Chemical compound NC(N)=NCCC[C@H](N)C(=O)N[C@@H](CCCN=C(N)N)C(=O)N[C@@H](CO)C(O)=O UISQLSIBJKEJSS-GUBZILKMSA-N 0.000 description 3
- ITVINTQUZMQWJR-QXEWZRGKSA-N Arg-Asn-Val Chemical compound [H]N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CC(N)=O)C(=O)N[C@@H](C(C)C)C(O)=O ITVINTQUZMQWJR-QXEWZRGKSA-N 0.000 description 3
- MFAMTAVAFBPXDC-LPEHRKFASA-N Arg-Asp-Pro Chemical compound C1C[C@@H](N(C1)C(=O)[C@H](CC(=O)O)NC(=O)[C@H](CCCN=C(N)N)N)C(=O)O MFAMTAVAFBPXDC-LPEHRKFASA-N 0.000 description 3
- GIVWETPOBCRTND-DCAQKATOSA-N Arg-Gln-Arg Chemical compound NC(N)=NCCC[C@H](N)C(=O)N[C@@H](CCC(N)=O)C(=O)N[C@@H](CCCN=C(N)N)C(O)=O GIVWETPOBCRTND-DCAQKATOSA-N 0.000 description 3
- VDBKFYYIBLXEIF-GUBZILKMSA-N Arg-Gln-Glu Chemical compound [H]N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CCC(N)=O)C(=O)N[C@@H](CCC(O)=O)C(O)=O VDBKFYYIBLXEIF-GUBZILKMSA-N 0.000 description 3
- BEXGZLUHRXTZCC-CIUDSAMLSA-N Arg-Gln-Ser Chemical compound C(C[C@@H](C(=O)N[C@@H](CCC(=O)N)C(=O)N[C@@H](CO)C(=O)O)N)CN=C(N)N BEXGZLUHRXTZCC-CIUDSAMLSA-N 0.000 description 3
- XLWSGICNBZGYTA-CIUDSAMLSA-N Arg-Glu-Asp Chemical compound NC(N)=NCCC[C@H](N)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CC(O)=O)C(O)=O XLWSGICNBZGYTA-CIUDSAMLSA-N 0.000 description 3
- ZATRYQNPUHGXCU-DTWKUNHWSA-N Arg-Gly-Pro Chemical compound C1C[C@@H](N(C1)C(=O)CNC(=O)[C@H](CCCN=C(N)N)N)C(=O)O ZATRYQNPUHGXCU-DTWKUNHWSA-N 0.000 description 3
- NKNILFJYKKHBKE-WPRPVWTQSA-N Arg-Gly-Val Chemical compound [H]N[C@@H](CCCNC(N)=N)C(=O)NCC(=O)N[C@@H](C(C)C)C(O)=O NKNILFJYKKHBKE-WPRPVWTQSA-N 0.000 description 3
- OFIYLHVAAJYRBC-HJWJTTGWSA-N Arg-Ile-Phe Chemical compound CC[C@H](C)[C@H](NC(=O)[C@@H](N)CCCNC(N)=N)C(=O)N[C@@H](Cc1ccccc1)C(O)=O OFIYLHVAAJYRBC-HJWJTTGWSA-N 0.000 description 3
- RIIVUOJDDQXHRV-SRVKXCTJSA-N Arg-Lys-Gln Chemical compound [H]N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CCCCN)C(=O)N[C@@H](CCC(N)=O)C(O)=O RIIVUOJDDQXHRV-SRVKXCTJSA-N 0.000 description 3
- CVXXSWQORBZAAA-SRVKXCTJSA-N Arg-Lys-Glu Chemical compound OC(=O)CC[C@@H](C(O)=O)NC(=O)[C@H](CCCCN)NC(=O)[C@@H](N)CCCN=C(N)N CVXXSWQORBZAAA-SRVKXCTJSA-N 0.000 description 3
- DPLFNLDACGGBAK-KKUMJFAQSA-N Arg-Phe-Glu Chemical compound C1=CC=C(C=C1)C[C@@H](C(=O)N[C@@H](CCC(=O)O)C(=O)O)NC(=O)[C@H](CCCN=C(N)N)N DPLFNLDACGGBAK-KKUMJFAQSA-N 0.000 description 3
- WKPXXXUSUHAXDE-SRVKXCTJSA-N Arg-Pro-Arg Chemical compound NC(N)=NCCC[C@H](N)C(=O)N1CCC[C@H]1C(=O)N[C@@H](CCCN=C(N)N)C(O)=O WKPXXXUSUHAXDE-SRVKXCTJSA-N 0.000 description 3
- NGYHSXDNNOFHNE-AVGNSLFASA-N Arg-Pro-Leu Chemical compound [H]N[C@@H](CCCNC(N)=N)C(=O)N1CCC[C@H]1C(=O)N[C@@H](CC(C)C)C(O)=O NGYHSXDNNOFHNE-AVGNSLFASA-N 0.000 description 3
- JOTRDIXZHNQYGP-DCAQKATOSA-N Arg-Ser-Lys Chemical compound C(CCN)C[C@@H](C(=O)O)NC(=O)[C@H](CO)NC(=O)[C@H](CCCN=C(N)N)N JOTRDIXZHNQYGP-DCAQKATOSA-N 0.000 description 3
- JPAWCMXVNZPJLO-IHRRRGAJSA-N Arg-Ser-Phe Chemical compound [H]N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CO)C(=O)N[C@@H](CC1=CC=CC=C1)C(O)=O JPAWCMXVNZPJLO-IHRRRGAJSA-N 0.000 description 3
- XVAPVJNJGLWGCS-ACZMJKKPSA-N Asn-Glu-Asn Chemical compound C(CC(=O)O)[C@@H](C(=O)N[C@@H](CC(=O)N)C(=O)O)NC(=O)[C@H](CC(=O)N)N XVAPVJNJGLWGCS-ACZMJKKPSA-N 0.000 description 3
- IBLAOXSULLECQZ-IUKAMOBKSA-N Asn-Ile-Thr Chemical compound C[C@@H](O)[C@@H](C(O)=O)NC(=O)[C@H]([C@@H](C)CC)NC(=O)[C@@H](N)CC(N)=O IBLAOXSULLECQZ-IUKAMOBKSA-N 0.000 description 3
- NLRJGXZWTKXRHP-DCAQKATOSA-N Asn-Leu-Arg Chemical compound [H]N[C@@H](CC(N)=O)C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](CCCNC(N)=N)C(O)=O NLRJGXZWTKXRHP-DCAQKATOSA-N 0.000 description 3
- HFPXZWPUVFVNLL-GUBZILKMSA-N Asn-Leu-Gln Chemical compound [H]N[C@@H](CC(N)=O)C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](CCC(N)=O)C(O)=O HFPXZWPUVFVNLL-GUBZILKMSA-N 0.000 description 3
- YUOXLJYVSZYPBJ-CIUDSAMLSA-N Asn-Pro-Glu Chemical compound [H]N[C@@H](CC(N)=O)C(=O)N1CCC[C@H]1C(=O)N[C@@H](CCC(O)=O)C(O)=O YUOXLJYVSZYPBJ-CIUDSAMLSA-N 0.000 description 3
- WSWYMRLTJVKRCE-ZLUOBGJFSA-N Asp-Ala-Asp Chemical compound OC(=O)C[C@H](N)C(=O)N[C@@H](C)C(=O)N[C@@H](CC(O)=O)C(O)=O WSWYMRLTJVKRCE-ZLUOBGJFSA-N 0.000 description 3
- OBMZMSLWNNWEJA-XNCRXQDQSA-N C1=CC=2C(C[C@@H]3NC(=O)[C@@H](NC(=O)[C@H](NC(=O)N(CC#CCN(CCCC[C@H](NC(=O)[C@@H](CC4=CC=CC=C4)NC3=O)C(=O)N)CC=C)NC(=O)[C@@H](N)C)CC3=CNC4=C3C=CC=C4)C)=CNC=2C=C1 Chemical compound C1=CC=2C(C[C@@H]3NC(=O)[C@@H](NC(=O)[C@H](NC(=O)N(CC#CCN(CCCC[C@H](NC(=O)[C@@H](CC4=CC=CC=C4)NC3=O)C(=O)N)CC=C)NC(=O)[C@@H](N)C)CC3=CNC4=C3C=CC=C4)C)=CNC=2C=C1 OBMZMSLWNNWEJA-XNCRXQDQSA-N 0.000 description 3
- 101100189913 Caenorhabditis elegans pept-1 gene Proteins 0.000 description 3
- 108020004705 Codon Proteins 0.000 description 3
- VKAWJBQTFCBHQY-GUBZILKMSA-N Cys-Gln-Lys Chemical compound C(CCN)C[C@@H](C(=O)O)NC(=O)[C@H](CCC(=O)N)NC(=O)[C@H](CS)N VKAWJBQTFCBHQY-GUBZILKMSA-N 0.000 description 3
- BNCKELUXXUYRNY-GUBZILKMSA-N Cys-Lys-Glu Chemical compound C(CCN)C[C@@H](C(=O)N[C@@H](CCC(=O)O)C(=O)O)NC(=O)[C@H](CS)N BNCKELUXXUYRNY-GUBZILKMSA-N 0.000 description 3
- UBHPUQAWSSNQLQ-DCAQKATOSA-N Cys-Pro-His Chemical compound C1C[C@H](N(C1)C(=O)[C@H](CS)N)C(=O)N[C@@H](CC2=CN=CN2)C(=O)O UBHPUQAWSSNQLQ-DCAQKATOSA-N 0.000 description 3
- 241000223221 Fusarium oxysporum Species 0.000 description 3
- LTLXPHKSQQILNF-CIUDSAMLSA-N Gln-Arg-Cys Chemical compound C(C[C@@H](C(=O)N[C@@H](CS)C(=O)O)NC(=O)[C@H](CCC(=O)N)N)CN=C(N)N LTLXPHKSQQILNF-CIUDSAMLSA-N 0.000 description 3
- SOBBAYVQSNXYPQ-ACZMJKKPSA-N Gln-Asn-Asn Chemical compound [H]N[C@@H](CCC(N)=O)C(=O)N[C@@H](CC(N)=O)C(=O)N[C@@H](CC(N)=O)C(O)=O SOBBAYVQSNXYPQ-ACZMJKKPSA-N 0.000 description 3
- ZPDVKYLJTOFQJV-WDSKDSINSA-N Gln-Asn-Gly Chemical compound [H]N[C@@H](CCC(N)=O)C(=O)N[C@@H](CC(N)=O)C(=O)NCC(O)=O ZPDVKYLJTOFQJV-WDSKDSINSA-N 0.000 description 3
- APWLZZSLCXLDCF-CIUDSAMLSA-N Gln-Cys-Met Chemical compound [H]N[C@@H](CCC(N)=O)C(=O)N[C@@H](CS)C(=O)N[C@@H](CCSC)C(O)=O APWLZZSLCXLDCF-CIUDSAMLSA-N 0.000 description 3
- RRBLZNIIMHSHQF-FXQIFTODSA-N Gln-Gln-Cys Chemical compound C(CC(=O)N)[C@@H](C(=O)N[C@@H](CCC(=O)N)C(=O)N[C@@H](CS)C(=O)O)N RRBLZNIIMHSHQF-FXQIFTODSA-N 0.000 description 3
- MAGNEQBFSBREJL-DCAQKATOSA-N Gln-Glu-Lys Chemical compound C(CCN)C[C@@H](C(=O)O)NC(=O)[C@H](CCC(=O)O)NC(=O)[C@H](CCC(=O)N)N MAGNEQBFSBREJL-DCAQKATOSA-N 0.000 description 3
- CLPQUWHBWXFJOX-BQBZGAKWSA-N Gln-Gly-Gln Chemical compound [H]N[C@@H](CCC(N)=O)C(=O)NCC(=O)N[C@@H](CCC(N)=O)C(O)=O CLPQUWHBWXFJOX-BQBZGAKWSA-N 0.000 description 3
- LGIKBBLQVSWUGK-DCAQKATOSA-N Gln-Leu-Gln Chemical compound [H]N[C@@H](CCC(N)=O)C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](CCC(N)=O)C(O)=O LGIKBBLQVSWUGK-DCAQKATOSA-N 0.000 description 3
- QKCZZAZNMMVICF-DCAQKATOSA-N Gln-Leu-Glu Chemical compound NC(=O)CC[C@H](N)C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](CCC(O)=O)C(O)=O QKCZZAZNMMVICF-DCAQKATOSA-N 0.000 description 3
- LGWNISYVKDNJRP-FXQIFTODSA-N Gln-Ser-Gln Chemical compound [H]N[C@@H](CCC(N)=O)C(=O)N[C@@H](CO)C(=O)N[C@@H](CCC(N)=O)C(O)=O LGWNISYVKDNJRP-FXQIFTODSA-N 0.000 description 3
- YLABFXCRQQMMHS-AVGNSLFASA-N Gln-Tyr-Cys Chemical compound C1=CC(=CC=C1C[C@@H](C(=O)N[C@@H](CS)C(=O)O)NC(=O)[C@H](CCC(=O)N)N)O YLABFXCRQQMMHS-AVGNSLFASA-N 0.000 description 3
- XXCDTYBVGMPIOA-FXQIFTODSA-N Glu-Asp-Glu Chemical compound OC(=O)CC[C@H](N)C(=O)N[C@@H](CC(O)=O)C(=O)N[C@@H](CCC(O)=O)C(O)=O XXCDTYBVGMPIOA-FXQIFTODSA-N 0.000 description 3
- CYHBMLHCQXXCCT-AVGNSLFASA-N Glu-Asp-Tyr Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N[C@@H](CC(O)=O)C(=O)N[C@@H](CC1=CC=C(O)C=C1)C(O)=O CYHBMLHCQXXCCT-AVGNSLFASA-N 0.000 description 3
- VNCNWQPIQYAMAK-ACZMJKKPSA-N Glu-Ser-Ser Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N[C@@H](CO)C(=O)N[C@@H](CO)C(O)=O VNCNWQPIQYAMAK-ACZMJKKPSA-N 0.000 description 3
- RQZGFWKQLPJOEQ-YUMQZZPRSA-N Gly-Arg-Gln Chemical compound C(C[C@@H](C(=O)N[C@@H](CCC(=O)N)C(=O)O)NC(=O)CN)CN=C(N)N RQZGFWKQLPJOEQ-YUMQZZPRSA-N 0.000 description 3
- OGCIHJPYKVSMTE-YUMQZZPRSA-N Gly-Arg-Glu Chemical compound [H]NCC(=O)N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CCC(O)=O)C(O)=O OGCIHJPYKVSMTE-YUMQZZPRSA-N 0.000 description 3
- MOJKRXIRAZPZLW-WDSKDSINSA-N Gly-Glu-Ala Chemical compound [H]NCC(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](C)C(O)=O MOJKRXIRAZPZLW-WDSKDSINSA-N 0.000 description 3
- HFXJIZNEXNIZIJ-BQBZGAKWSA-N Gly-Glu-Gln Chemical compound NCC(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CCC(N)=O)C(O)=O HFXJIZNEXNIZIJ-BQBZGAKWSA-N 0.000 description 3
- HQRHFUYMGCHHJS-LURJTMIESA-N Gly-Gly-Arg Chemical compound NCC(=O)NCC(=O)N[C@H](C(O)=O)CCCN=C(N)N HQRHFUYMGCHHJS-LURJTMIESA-N 0.000 description 3
- TVTZEOHWHUVYCG-KYNKHSRBSA-N Gly-Thr-Thr Chemical compound [H]NCC(=O)N[C@@H]([C@@H](C)O)C(=O)N[C@@H]([C@@H](C)O)C(O)=O TVTZEOHWHUVYCG-KYNKHSRBSA-N 0.000 description 3
- JBCLFWXMTIKCCB-UHFFFAOYSA-N H-Gly-Phe-OH Natural products NCC(=O)NC(C(O)=O)CC1=CC=CC=C1 JBCLFWXMTIKCCB-UHFFFAOYSA-N 0.000 description 3
- ZNNNYCXPCKACHX-DCAQKATOSA-N His-Gln-Gln Chemical compound [H]N[C@@H](CC1=CNC=N1)C(=O)N[C@@H](CCC(N)=O)C(=O)N[C@@H](CCC(N)=O)C(O)=O ZNNNYCXPCKACHX-DCAQKATOSA-N 0.000 description 3
- TVRMJKNELJKNRS-GUBZILKMSA-N His-Glu-Asn Chemical compound C1=C(NC=N1)C[C@@H](C(=O)N[C@@H](CCC(=O)O)C(=O)N[C@@H](CC(=O)N)C(=O)O)N TVRMJKNELJKNRS-GUBZILKMSA-N 0.000 description 3
- PYNUBZSXKQKAHL-UWVGGRQHSA-N His-Gly-Arg Chemical compound [H]N[C@@H](CC1=CNC=N1)C(=O)NCC(=O)N[C@@H](CCCNC(N)=N)C(O)=O PYNUBZSXKQKAHL-UWVGGRQHSA-N 0.000 description 3
- ORERHHPZDDEMSC-VGDYDELISA-N His-Ile-Ser Chemical compound CC[C@H](C)[C@@H](C(=O)N[C@@H](CO)C(=O)O)NC(=O)[C@H](CC1=CN=CN1)N ORERHHPZDDEMSC-VGDYDELISA-N 0.000 description 3
- XVZJRZQIHJMUBG-TUBUOCAGSA-N His-Thr-Ile Chemical compound CC[C@H](C)[C@@H](C(=O)O)NC(=O)[C@H]([C@@H](C)O)NC(=O)[C@H](CC1=CN=CN1)N XVZJRZQIHJMUBG-TUBUOCAGSA-N 0.000 description 3
- RWIKBYVJQAJYDP-BJDJZHNGSA-N Ile-Ala-Lys Chemical compound CC[C@H](C)[C@H](N)C(=O)N[C@@H](C)C(=O)N[C@H](C(O)=O)CCCCN RWIKBYVJQAJYDP-BJDJZHNGSA-N 0.000 description 3
- FVEWRQXNISSYFO-ZPFDUUQYSA-N Ile-Arg-Glu Chemical compound CC[C@H](C)[C@@H](C(=O)N[C@@H](CCCN=C(N)N)C(=O)N[C@@H](CCC(=O)O)C(=O)O)N FVEWRQXNISSYFO-ZPFDUUQYSA-N 0.000 description 3
- QADCTXFNLZBZAB-GHCJXIJMSA-N Ile-Asn-Ala Chemical compound CC[C@H](C)[C@@H](C(=O)N[C@@H](CC(=O)N)C(=O)N[C@@H](C)C(=O)O)N QADCTXFNLZBZAB-GHCJXIJMSA-N 0.000 description 3
- LDRALPZEVHVXEK-KBIXCLLPSA-N Ile-Cys-Glu Chemical compound CC[C@H](C)[C@@H](C(=O)N[C@@H](CS)C(=O)N[C@@H](CCC(=O)O)C(=O)O)N LDRALPZEVHVXEK-KBIXCLLPSA-N 0.000 description 3
- BBQABUDWDUKJMB-LZXPERKUSA-N Ile-Ile-Ile Chemical compound CC[C@H](C)[C@H]([NH3+])C(=O)N[C@@H]([C@@H](C)CC)C(=O)N[C@@H]([C@@H](C)CC)C([O-])=O BBQABUDWDUKJMB-LZXPERKUSA-N 0.000 description 3
- XLXPYSDGMXTTNQ-UHFFFAOYSA-N Ile-Phe-Leu Natural products CCC(C)C(N)C(=O)NC(C(=O)NC(CC(C)C)C(O)=O)CC1=CC=CC=C1 XLXPYSDGMXTTNQ-UHFFFAOYSA-N 0.000 description 3
- 108010065920 Insulin Lispro Proteins 0.000 description 3
- KFKWRHQBZQICHA-STQMWFEESA-N L-leucyl-L-phenylalanine Natural products CC(C)C[C@H](N)C(=O)N[C@H](C(O)=O)CC1=CC=CC=C1 KFKWRHQBZQICHA-STQMWFEESA-N 0.000 description 3
- KAFOIVJDVSZUMD-DCAQKATOSA-N Leu-Gln-Gln Chemical compound [H]N[C@@H](CC(C)C)C(=O)N[C@@H](CCC(N)=O)C(=O)N[C@@H](CCC(N)=O)C(O)=O KAFOIVJDVSZUMD-DCAQKATOSA-N 0.000 description 3
- PPQRKXHCLYCBSP-IHRRRGAJSA-N Leu-Leu-Met Chemical compound CC(C)C[C@@H](C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](CCSC)C(=O)O)N PPQRKXHCLYCBSP-IHRRRGAJSA-N 0.000 description 3
- BIZNDKMFQHDOIE-KKUMJFAQSA-N Leu-Phe-Asn Chemical compound CC(C)C[C@H](N)C(=O)N[C@H](C(=O)N[C@@H](CC(N)=O)C(O)=O)CC1=CC=CC=C1 BIZNDKMFQHDOIE-KKUMJFAQSA-N 0.000 description 3
- AKVBOOKXVAMKSS-GUBZILKMSA-N Leu-Ser-Gln Chemical compound [H]N[C@@H](CC(C)C)C(=O)N[C@@H](CO)C(=O)N[C@@H](CCC(N)=O)C(O)=O AKVBOOKXVAMKSS-GUBZILKMSA-N 0.000 description 3
- IWMJFLJQHIDZQW-KKUMJFAQSA-N Leu-Ser-Phe Chemical compound CC(C)C[C@H](N)C(=O)N[C@@H](CO)C(=O)N[C@H](C(O)=O)CC1=CC=CC=C1 IWMJFLJQHIDZQW-KKUMJFAQSA-N 0.000 description 3
- GQUDMNDPQTXZRV-DCAQKATOSA-N Lys-Arg-Asp Chemical compound [H]N[C@@H](CCCCN)C(=O)N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CC(O)=O)C(O)=O GQUDMNDPQTXZRV-DCAQKATOSA-N 0.000 description 3
- PBIPLDMFHAICIP-DCAQKATOSA-N Lys-Glu-Glu Chemical compound [H]N[C@@H](CCCCN)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CCC(O)=O)C(O)=O PBIPLDMFHAICIP-DCAQKATOSA-N 0.000 description 3
- IMAKMJCBYCSMHM-AVGNSLFASA-N Lys-Glu-Lys Chemical compound NCCCC[C@H](N)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@H](C(O)=O)CCCCN IMAKMJCBYCSMHM-AVGNSLFASA-N 0.000 description 3
- ZXFRGTAIIZHNHG-AJNGGQMLSA-N Lys-Ile-Leu Chemical compound CC[C@H](C)[C@@H](C(=O)N[C@@H](CC(C)C)C(=O)O)NC(=O)[C@H](CCCCN)N ZXFRGTAIIZHNHG-AJNGGQMLSA-N 0.000 description 3
- PWPBGAJJYJJVPI-PJODQICGSA-N Met-Ala-Trp Chemical compound C1=CC=C2C(C[C@H](NC(=O)[C@H](C)NC(=O)[C@@H](N)CCSC)C(O)=O)=CNC2=C1 PWPBGAJJYJJVPI-PJODQICGSA-N 0.000 description 3
- NLHSFJQUHGCWSD-PYJNHQTQSA-N Met-Ile-His Chemical compound N[C@@H](CCSC)C(=O)N[C@@H]([C@@H](C)CC)C(=O)N[C@@H](CC1=CNC=N1)C(O)=O NLHSFJQUHGCWSD-PYJNHQTQSA-N 0.000 description 3
- 108010047562 NGR peptide Proteins 0.000 description 3
- 101100342977 Neurospora crassa (strain ATCC 24698 / 74-OR23-1A / CBS 708.71 / DSM 1257 / FGSC 987) leu-1 gene Proteins 0.000 description 3
- 241000283973 Oryctolagus cuniculus Species 0.000 description 3
- 101710176384 Peptide 1 Proteins 0.000 description 3
- KAHUBGWSIQNZQQ-KKUMJFAQSA-N Phe-Asn-Lys Chemical compound NCCCC[C@@H](C(O)=O)NC(=O)[C@H](CC(N)=O)NC(=O)[C@@H](N)CC1=CC=CC=C1 KAHUBGWSIQNZQQ-KKUMJFAQSA-N 0.000 description 3
- VZFPYFRVHMSSNA-JURCDPSOSA-N Phe-Ile-Ala Chemical compound OC(=O)[C@H](C)NC(=O)[C@H]([C@@H](C)CC)NC(=O)[C@@H](N)CC1=CC=CC=C1 VZFPYFRVHMSSNA-JURCDPSOSA-N 0.000 description 3
- KXUZHWXENMYOHC-QEJZJMRPSA-N Phe-Leu-Ala Chemical compound [H]N[C@@H](CC1=CC=CC=C1)C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](C)C(O)=O KXUZHWXENMYOHC-QEJZJMRPSA-N 0.000 description 3
- RSPUIENXSJYZQO-JYJNAYRXSA-N Phe-Leu-Gln Chemical compound NC(=O)CC[C@@H](C(O)=O)NC(=O)[C@H](CC(C)C)NC(=O)[C@@H](N)CC1=CC=CC=C1 RSPUIENXSJYZQO-JYJNAYRXSA-N 0.000 description 3
- YTILBRIUASDGBL-BZSNNMDCSA-N Phe-Leu-Leu Chemical compound CC(C)C[C@@H](C(O)=O)NC(=O)[C@H](CC(C)C)NC(=O)[C@@H](N)CC1=CC=CC=C1 YTILBRIUASDGBL-BZSNNMDCSA-N 0.000 description 3
- MMJJFXWMCMJMQA-STQMWFEESA-N Phe-Pro-Gly Chemical compound C([C@H](N)C(=O)N1[C@@H](CCC1)C(=O)NCC(O)=O)C1=CC=CC=C1 MMJJFXWMCMJMQA-STQMWFEESA-N 0.000 description 3
- ZYNBEWGJFXTBDU-ACRUOGEOSA-N Phe-Tyr-Leu Chemical compound CC(C)C[C@@H](C(=O)O)NC(=O)[C@H](CC1=CC=C(C=C1)O)NC(=O)[C@H](CC2=CC=CC=C2)N ZYNBEWGJFXTBDU-ACRUOGEOSA-N 0.000 description 3
- VIIRRNQMMIHYHQ-XHSDSOJGSA-N Phe-Val-Pro Chemical compound CC(C)[C@@H](C(=O)N1CCC[C@@H]1C(=O)O)NC(=O)[C@H](CC2=CC=CC=C2)N VIIRRNQMMIHYHQ-XHSDSOJGSA-N 0.000 description 3
- QBFONMUYNSNKIX-AVGNSLFASA-N Pro-Arg-His Chemical compound C1C[C@H](NC1)C(=O)N[C@@H](CCCN=C(N)N)C(=O)N[C@@H](CC2=CN=CN2)C(=O)O QBFONMUYNSNKIX-AVGNSLFASA-N 0.000 description 3
- JMVQDLDPDBXAAX-YUMQZZPRSA-N Pro-Gly-Gln Chemical compound NC(=O)CC[C@@H](C(O)=O)NC(=O)CNC(=O)[C@@H]1CCCN1 JMVQDLDPDBXAAX-YUMQZZPRSA-N 0.000 description 3
- YHUBAXGAAYULJY-ULQDDVLXSA-N Pro-Tyr-Leu Chemical compound [H]N1CCC[C@H]1C(=O)N[C@@H](CC1=CC=C(O)C=C1)C(=O)N[C@@H](CC(C)C)C(O)=O YHUBAXGAAYULJY-ULQDDVLXSA-N 0.000 description 3
- 241000208465 Proteaceae Species 0.000 description 3
- WDXYVIIVDIDOSX-DCAQKATOSA-N Ser-Arg-Leu Chemical compound CC(C)C[C@@H](C(O)=O)NC(=O)[C@@H](NC(=O)[C@@H](N)CO)CCCN=C(N)N WDXYVIIVDIDOSX-DCAQKATOSA-N 0.000 description 3
- UOLGINIHBRIECN-FXQIFTODSA-N Ser-Glu-Glu Chemical compound [H]N[C@@H](CO)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CCC(O)=O)C(O)=O UOLGINIHBRIECN-FXQIFTODSA-N 0.000 description 3
- RIAKPZVSNBBNRE-BJDJZHNGSA-N Ser-Ile-Leu Chemical compound OC[C@H](N)C(=O)N[C@@H]([C@@H](C)CC)C(=O)N[C@@H](CC(C)C)C(O)=O RIAKPZVSNBBNRE-BJDJZHNGSA-N 0.000 description 3
- GZSZPKSBVAOGIE-CIUDSAMLSA-N Ser-Lys-Ala Chemical compound [H]N[C@@H](CO)C(=O)N[C@@H](CCCCN)C(=O)N[C@@H](C)C(O)=O GZSZPKSBVAOGIE-CIUDSAMLSA-N 0.000 description 3
- RXUOAOOZIWABBW-XGEHTFHBSA-N Ser-Thr-Arg Chemical compound OC[C@H](N)C(=O)N[C@@H]([C@H](O)C)C(=O)N[C@H](C(O)=O)CCCN=C(N)N RXUOAOOZIWABBW-XGEHTFHBSA-N 0.000 description 3
- PCJLFYBAQZQOFE-KATARQTJSA-N Ser-Thr-Lys Chemical compound C[C@H]([C@@H](C(=O)N[C@@H](CCCCN)C(=O)O)NC(=O)[C@H](CO)N)O PCJLFYBAQZQOFE-KATARQTJSA-N 0.000 description 3
- JGUWRQWULDWNCM-FXQIFTODSA-N Ser-Val-Ser Chemical compound [H]N[C@@H](CO)C(=O)N[C@@H](C(C)C)C(=O)N[C@@H](CO)C(O)=O JGUWRQWULDWNCM-FXQIFTODSA-N 0.000 description 3
- KBLYJPQSNGTDIU-LOKLDPHHSA-N Thr-Glu-Pro Chemical compound C[C@H]([C@@H](C(=O)N[C@@H](CCC(=O)O)C(=O)N1CCC[C@@H]1C(=O)O)N)O KBLYJPQSNGTDIU-LOKLDPHHSA-N 0.000 description 3
- DXPURPNJDFCKKO-RHYQMDGZSA-N Thr-Lys-Val Chemical compound CC(C)[C@H](NC(=O)[C@H](CCCCN)NC(=O)[C@@H](N)[C@@H](C)O)C(O)=O DXPURPNJDFCKKO-RHYQMDGZSA-N 0.000 description 3
- HNIWONZFMIPCCT-SIXJUCDHSA-N Trp-His-Ile Chemical compound CC[C@H](C)[C@@H](C(=O)O)NC(=O)[C@H](CC1=CN=CN1)NC(=O)[C@H](CC2=CNC3=CC=CC=C32)N HNIWONZFMIPCCT-SIXJUCDHSA-N 0.000 description 3
- ADBDQGBDNUTRDB-ULQDDVLXSA-N Tyr-Arg-Leu Chemical compound [H]N[C@@H](CC1=CC=C(O)C=C1)C(=O)N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CC(C)C)C(O)=O ADBDQGBDNUTRDB-ULQDDVLXSA-N 0.000 description 3
- WAPFQMXRSDEGOE-IHRRRGAJSA-N Tyr-Glu-Gln Chemical compound [H]N[C@@H](CC1=CC=C(O)C=C1)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CCC(N)=O)C(O)=O WAPFQMXRSDEGOE-IHRRRGAJSA-N 0.000 description 3
- SLCSPPCQWUHPPO-JYJNAYRXSA-N Tyr-Glu-Lys Chemical compound NCCCC[C@@H](C(O)=O)NC(=O)[C@H](CCC(O)=O)NC(=O)[C@@H](N)CC1=CC=C(O)C=C1 SLCSPPCQWUHPPO-JYJNAYRXSA-N 0.000 description 3
- JLKVWTICWVWGSK-JYJNAYRXSA-N Tyr-Lys-Glu Chemical compound OC(=O)CC[C@@H](C(O)=O)NC(=O)[C@H](CCCCN)NC(=O)[C@@H](N)CC1=CC=C(O)C=C1 JLKVWTICWVWGSK-JYJNAYRXSA-N 0.000 description 3
- SOAUMCDLIUGXJJ-SRVKXCTJSA-N Tyr-Ser-Asn Chemical compound [H]N[C@@H](CC1=CC=C(O)C=C1)C(=O)N[C@@H](CO)C(=O)N[C@@H](CC(N)=O)C(O)=O SOAUMCDLIUGXJJ-SRVKXCTJSA-N 0.000 description 3
- COYSIHFOCOMGCF-UHFFFAOYSA-N Val-Arg-Gly Natural products CC(C)C(N)C(=O)NC(C(=O)NCC(O)=O)CCCN=C(N)N COYSIHFOCOMGCF-UHFFFAOYSA-N 0.000 description 3
- HGJRMXOWUWVUOA-GVXVVHGQSA-N Val-Leu-Gln Chemical compound CC(C)C[C@@H](C(=O)N[C@@H](CCC(=O)N)C(=O)O)NC(=O)[C@H](C(C)C)N HGJRMXOWUWVUOA-GVXVVHGQSA-N 0.000 description 3
- ZHQWPWQNVRCXAX-XQQFMLRXSA-N Val-Leu-Pro Chemical compound CC(C)C[C@@H](C(=O)N1CCC[C@@H]1C(=O)O)NC(=O)[C@H](C(C)C)N ZHQWPWQNVRCXAX-XQQFMLRXSA-N 0.000 description 3
- MJOUSKQHAIARKI-JYJNAYRXSA-N Val-Phe-Val Chemical compound CC(C)[C@H](N)C(=O)N[C@H](C(=O)N[C@@H](C(C)C)C(O)=O)CC1=CC=CC=C1 MJOUSKQHAIARKI-JYJNAYRXSA-N 0.000 description 3
- NZYNRRGJJVSSTJ-GUBZILKMSA-N Val-Ser-Val Chemical compound CC(C)[C@H](N)C(=O)N[C@@H](CO)C(=O)N[C@@H](C(C)C)C(O)=O NZYNRRGJJVSSTJ-GUBZILKMSA-N 0.000 description 3
- RTJPAGFXOWEBAI-SRVKXCTJSA-N Val-Val-Arg Chemical compound CC(C)[C@H](N)C(=O)N[C@@H](C(C)C)C(=O)N[C@H](C(O)=O)CCCN=C(N)N RTJPAGFXOWEBAI-SRVKXCTJSA-N 0.000 description 3
- 241001123668 Verticillium dahliae Species 0.000 description 3
- 125000003295 alanine group Chemical group N[C@@H](C)C(=O)* 0.000 description 3
- 108010076324 alanyl-glycyl-glycine Proteins 0.000 description 3
- 108010086434 alanyl-seryl-glycine Proteins 0.000 description 3
- BFNBIHQBYMNNAN-UHFFFAOYSA-N ammonium sulfate Chemical compound N.N.OS(O)(=O)=O BFNBIHQBYMNNAN-UHFFFAOYSA-N 0.000 description 3
- 229910052921 ammonium sulfate Inorganic materials 0.000 description 3
- 239000001166 ammonium sulphate Substances 0.000 description 3
- 235000011130 ammonium sulphate Nutrition 0.000 description 3
- 229940121375 antifungal agent Drugs 0.000 description 3
- 108010072041 arginyl-glycyl-aspartic acid Proteins 0.000 description 3
- 108010069205 aspartyl-phenylalanine Proteins 0.000 description 3
- 108010092854 aspartyllysine Proteins 0.000 description 3
- 238000010276 construction Methods 0.000 description 3
- 108010060199 cysteinylproline Proteins 0.000 description 3
- 108010009297 diglycyl-histidine Proteins 0.000 description 3
- 239000013604 expression vector Substances 0.000 description 3
- 108010013768 glutamyl-aspartyl-proline Proteins 0.000 description 3
- XBGGUPMXALFZOT-UHFFFAOYSA-N glycyl-L-tyrosine hemihydrate Natural products NCC(=O)NC(C(O)=O)CC1=CC=C(O)C=C1 XBGGUPMXALFZOT-UHFFFAOYSA-N 0.000 description 3
- 108010062266 glycyl-glycyl-argininal Proteins 0.000 description 3
- 108010089804 glycyl-threonine Proteins 0.000 description 3
- 108010048994 glycyl-tyrosyl-alanine Proteins 0.000 description 3
- 108010037850 glycylvaline Proteins 0.000 description 3
- 229960004198 guanidine Drugs 0.000 description 3
- PJJJBBJSCAKJQF-UHFFFAOYSA-N guanidinium chloride Chemical compound [Cl-].NC(N)=[NH2+] PJJJBBJSCAKJQF-UHFFFAOYSA-N 0.000 description 3
- 238000004128 high performance liquid chromatography Methods 0.000 description 3
- 108010092114 histidylphenylalanine Proteins 0.000 description 3
- 230000006698 induction Effects 0.000 description 3
- 208000015181 infectious disease Diseases 0.000 description 3
- 108010038320 lysylphenylalanine Proteins 0.000 description 3
- 238000005259 measurement Methods 0.000 description 3
- 108020004999 messenger RNA Proteins 0.000 description 3
- 229910052757 nitrogen Inorganic materials 0.000 description 3
- 239000002245 particle Substances 0.000 description 3
- 230000001717 pathogenic effect Effects 0.000 description 3
- 238000002360 preparation method Methods 0.000 description 3
- 125000001500 prolyl group Chemical group [H]N1C([H])(C(=O)[*])C([H])([H])C([H])([H])C1([H])[H] 0.000 description 3
- 108010014614 prolyl-glycyl-proline Proteins 0.000 description 3
- 108010070643 prolylglutamic acid Proteins 0.000 description 3
- 230000002829 reductive effect Effects 0.000 description 3
- 108010026333 seryl-proline Proteins 0.000 description 3
- 239000007787 solid Substances 0.000 description 3
- 239000000243 solution Substances 0.000 description 3
- 238000004611 spectroscopical analysis Methods 0.000 description 3
- 108010005652 splenotritin Proteins 0.000 description 3
- 238000003756 stirring Methods 0.000 description 3
- 125000001493 tyrosinyl group Chemical group [H]OC1=C([H])C([H])=C(C([H])=C1[H])C([H])([H])C([H])(N([H])[H])C(*)=O 0.000 description 3
- 108010077037 tyrosyl-tyrosyl-phenylalanine Proteins 0.000 description 3
- YBJHBAHKTGYVGT-ZKWXMUAHSA-N (+)-Biotin Chemical compound N1C(=O)N[C@@H]2[C@H](CCCCC(=O)O)SC[C@@H]21 YBJHBAHKTGYVGT-ZKWXMUAHSA-N 0.000 description 2
- KFDVPJUYSDEJTH-UHFFFAOYSA-N 4-ethenylpyridine Chemical compound C=CC1=CC=NC=C1 KFDVPJUYSDEJTH-UHFFFAOYSA-N 0.000 description 2
- CXRCVCURMBFFOL-FXQIFTODSA-N Ala-Ala-Pro Chemical compound C[C@H](N)C(=O)N[C@@H](C)C(=O)N1CCC[C@H]1C(O)=O CXRCVCURMBFFOL-FXQIFTODSA-N 0.000 description 2
- ZEXDYVGDZJBRMO-ACZMJKKPSA-N Ala-Asn-Gln Chemical compound C[C@@H](C(=O)N[C@@H](CC(=O)N)C(=O)N[C@@H](CCC(=O)N)C(=O)O)N ZEXDYVGDZJBRMO-ACZMJKKPSA-N 0.000 description 2
- NHCPCLJZRSIDHS-ZLUOBGJFSA-N Ala-Asp-Ala Chemical compound [H]N[C@@H](C)C(=O)N[C@@H](CC(O)=O)C(=O)N[C@@H](C)C(O)=O NHCPCLJZRSIDHS-ZLUOBGJFSA-N 0.000 description 2
- WGDNWOMKBUXFHR-BQBZGAKWSA-N Ala-Gly-Arg Chemical compound C[C@H](N)C(=O)NCC(=O)N[C@H](C(O)=O)CCCN=C(N)N WGDNWOMKBUXFHR-BQBZGAKWSA-N 0.000 description 2
- IFKQPMZRDQZSHI-GHCJXIJMSA-N Ala-Ile-Asn Chemical compound [H]N[C@@H](C)C(=O)N[C@@H]([C@@H](C)CC)C(=O)N[C@@H](CC(N)=O)C(O)=O IFKQPMZRDQZSHI-GHCJXIJMSA-N 0.000 description 2
- TZDNWXDLYFIFPT-BJDJZHNGSA-N Ala-Ile-Leu Chemical compound [H]N[C@@H](C)C(=O)N[C@@H]([C@@H](C)CC)C(=O)N[C@@H](CC(C)C)C(O)=O TZDNWXDLYFIFPT-BJDJZHNGSA-N 0.000 description 2
- YHKANGMVQWRMAP-DCAQKATOSA-N Ala-Leu-Arg Chemical compound C[C@H](N)C(=O)N[C@@H](CC(C)C)C(=O)N[C@H](C(O)=O)CCCN=C(N)N YHKANGMVQWRMAP-DCAQKATOSA-N 0.000 description 2
- DWYROCSXOOMOEU-CIUDSAMLSA-N Ala-Met-Glu Chemical compound C[C@@H](C(=O)N[C@@H](CCSC)C(=O)N[C@@H](CCC(=O)O)C(=O)O)N DWYROCSXOOMOEU-CIUDSAMLSA-N 0.000 description 2
- OSRZOHXQCUFIQG-FPMFFAJLSA-N Ala-Phe-Pro Chemical compound C([C@H](NC(=O)[C@@H]([NH3+])C)C(=O)N1[C@H](CCC1)C([O-])=O)C1=CC=CC=C1 OSRZOHXQCUFIQG-FPMFFAJLSA-N 0.000 description 2
- YCRAFFCYWOUEOF-DLOVCJGASA-N Ala-Phe-Ser Chemical compound OC[C@@H](C(O)=O)NC(=O)[C@@H](NC(=O)[C@@H](N)C)CC1=CC=CC=C1 YCRAFFCYWOUEOF-DLOVCJGASA-N 0.000 description 2
- IHMCQESUJVZTKW-UBHSHLNASA-N Ala-Phe-Val Chemical compound CC(C)[C@@H](C(O)=O)NC(=O)[C@@H](NC(=O)[C@H](C)N)CC1=CC=CC=C1 IHMCQESUJVZTKW-UBHSHLNASA-N 0.000 description 2
- MMLHRUJLOUSRJX-CIUDSAMLSA-N Ala-Ser-Lys Chemical compound C[C@H](N)C(=O)N[C@@H](CO)C(=O)N[C@H](C(O)=O)CCCCN MMLHRUJLOUSRJX-CIUDSAMLSA-N 0.000 description 2
- NCQMBSJGJMYKCK-ZLUOBGJFSA-N Ala-Ser-Ser Chemical compound [H]N[C@@H](C)C(=O)N[C@@H](CO)C(=O)N[C@@H](CO)C(O)=O NCQMBSJGJMYKCK-ZLUOBGJFSA-N 0.000 description 2
- BGGAIXWIZCIFSG-XDTLVQLUSA-N Ala-Tyr-Glu Chemical compound [H]N[C@@H](C)C(=O)N[C@@H](CC1=CC=C(O)C=C1)C(=O)N[C@@H](CCC(O)=O)C(O)=O BGGAIXWIZCIFSG-XDTLVQLUSA-N 0.000 description 2
- DHONNEYAZPNGSG-UBHSHLNASA-N Ala-Val-Phe Chemical compound C[C@H](N)C(=O)N[C@@H](C(C)C)C(=O)N[C@H](C(O)=O)CC1=CC=CC=C1 DHONNEYAZPNGSG-UBHSHLNASA-N 0.000 description 2
- 241000223600 Alternaria Species 0.000 description 2
- SGYSTDWPNPKJPP-GUBZILKMSA-N Arg-Ala-Arg Chemical compound NC(=N)NCCC[C@H](N)C(=O)N[C@@H](C)C(=O)N[C@@H](CCCNC(N)=N)C(O)=O SGYSTDWPNPKJPP-GUBZILKMSA-N 0.000 description 2
- GIVATXIGCXFQQA-FXQIFTODSA-N Arg-Ala-Ser Chemical compound OC[C@@H](C(O)=O)NC(=O)[C@H](C)NC(=O)[C@@H](N)CCCN=C(N)N GIVATXIGCXFQQA-FXQIFTODSA-N 0.000 description 2
- OVVUNXXROOFSIM-SDDRHHMPSA-N Arg-Arg-Pro Chemical compound C1C[C@@H](N(C1)C(=O)[C@H](CCCN=C(N)N)NC(=O)[C@H](CCCN=C(N)N)N)C(=O)O OVVUNXXROOFSIM-SDDRHHMPSA-N 0.000 description 2
- PQWTZSNVWSOFFK-FXQIFTODSA-N Arg-Asp-Asn Chemical compound C(C[C@@H](C(=O)N[C@@H](CC(=O)O)C(=O)N[C@@H](CC(=O)N)C(=O)O)N)CN=C(N)N PQWTZSNVWSOFFK-FXQIFTODSA-N 0.000 description 2
- HJAICMSAKODKRF-GUBZILKMSA-N Arg-Cys-Arg Chemical compound NC(N)=NCCC[C@H](N)C(=O)N[C@@H](CS)C(=O)N[C@@H](CCCN=C(N)N)C(O)=O HJAICMSAKODKRF-GUBZILKMSA-N 0.000 description 2
- GDVDRMUYICMNFJ-CIUDSAMLSA-N Arg-Cys-Glu Chemical compound [H]N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CS)C(=O)N[C@@H](CCC(O)=O)C(O)=O GDVDRMUYICMNFJ-CIUDSAMLSA-N 0.000 description 2
- RWDVGVPHEWOZMO-GUBZILKMSA-N Arg-Cys-Val Chemical compound CC(C)[C@H](NC(=O)[C@H](CS)NC(=O)[C@@H](N)CCCNC(N)=N)C(O)=O RWDVGVPHEWOZMO-GUBZILKMSA-N 0.000 description 2
- LLZXKVAAEWBUPB-KKUMJFAQSA-N Arg-Gln-Phe Chemical compound [H]N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CCC(N)=O)C(=O)N[C@@H](CC1=CC=CC=C1)C(O)=O LLZXKVAAEWBUPB-KKUMJFAQSA-N 0.000 description 2
- PBSOQGZLPFVXPU-YUMQZZPRSA-N Arg-Glu-Gly Chemical compound [H]N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CCC(O)=O)C(=O)NCC(O)=O PBSOQGZLPFVXPU-YUMQZZPRSA-N 0.000 description 2
- JAYIQMNQDMOBFY-KKUMJFAQSA-N Arg-Glu-Tyr Chemical compound [H]N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CC1=CC=C(O)C=C1)C(O)=O JAYIQMNQDMOBFY-KKUMJFAQSA-N 0.000 description 2
- GOWZVQXTHUCNSQ-NHCYSSNCSA-N Arg-Glu-Val Chemical compound [H]N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](C(C)C)C(O)=O GOWZVQXTHUCNSQ-NHCYSSNCSA-N 0.000 description 2
- YBIAYFFIVAZXPK-AVGNSLFASA-N Arg-His-Arg Chemical compound [H]N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CC1=CNC=N1)C(=O)N[C@@H](CCCNC(N)=N)C(O)=O YBIAYFFIVAZXPK-AVGNSLFASA-N 0.000 description 2
- WMEVEPXNCMKNGH-IHRRRGAJSA-N Arg-Leu-His Chemical compound CC(C)C[C@@H](C(=O)N[C@@H](CC1=CN=CN1)C(=O)O)NC(=O)[C@H](CCCN=C(N)N)N WMEVEPXNCMKNGH-IHRRRGAJSA-N 0.000 description 2
- COXMUHNBYCVVRG-DCAQKATOSA-N Arg-Leu-Ser Chemical compound [H]N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](CO)C(O)=O COXMUHNBYCVVRG-DCAQKATOSA-N 0.000 description 2
- MNBHKGYCLBUIBC-UFYCRDLUSA-N Arg-Phe-Phe Chemical compound C([C@H](NC(=O)[C@H](CCCNC(N)=N)N)C(=O)N[C@@H](CC=1C=CC=CC=1)C(O)=O)C1=CC=CC=C1 MNBHKGYCLBUIBC-UFYCRDLUSA-N 0.000 description 2
- ATABBWFGOHKROJ-GUBZILKMSA-N Arg-Pro-Ser Chemical compound [H]N[C@@H](CCCNC(N)=N)C(=O)N1CCC[C@H]1C(=O)N[C@@H](CO)C(O)=O ATABBWFGOHKROJ-GUBZILKMSA-N 0.000 description 2
- ISJWBVIYRBAXEB-CIUDSAMLSA-N Arg-Ser-Glu Chemical compound [H]N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CO)C(=O)N[C@@H](CCC(O)=O)C(O)=O ISJWBVIYRBAXEB-CIUDSAMLSA-N 0.000 description 2
- DNLQVHBBMPZUGJ-BQBZGAKWSA-N Arg-Ser-Gly Chemical compound [H]N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CO)C(=O)NCC(O)=O DNLQVHBBMPZUGJ-BQBZGAKWSA-N 0.000 description 2
- KMFPQTITXUKJOV-DCAQKATOSA-N Arg-Ser-Leu Chemical compound [H]N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CO)C(=O)N[C@@H](CC(C)C)C(O)=O KMFPQTITXUKJOV-DCAQKATOSA-N 0.000 description 2
- 239000004475 Arginine Substances 0.000 description 2
- QQEWINYJRFBLNN-DLOVCJGASA-N Asn-Ala-Phe Chemical compound NC(=O)C[C@H](N)C(=O)N[C@@H](C)C(=O)N[C@H](C(O)=O)CC1=CC=CC=C1 QQEWINYJRFBLNN-DLOVCJGASA-N 0.000 description 2
- QEYJFBMTSMLPKZ-ZKWXMUAHSA-N Asn-Ala-Val Chemical compound [H]N[C@@H](CC(N)=O)C(=O)N[C@@H](C)C(=O)N[C@@H](C(C)C)C(O)=O QEYJFBMTSMLPKZ-ZKWXMUAHSA-N 0.000 description 2
- HOIFSHOLNKQCSA-FXQIFTODSA-N Asn-Arg-Asp Chemical compound [H]N[C@@H](CC(N)=O)C(=O)N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CC(O)=O)C(O)=O HOIFSHOLNKQCSA-FXQIFTODSA-N 0.000 description 2
- LJUOLNXOWSWGKF-ACZMJKKPSA-N Asn-Asn-Glu Chemical compound C(CC(=O)O)[C@@H](C(=O)O)NC(=O)[C@H](CC(=O)N)NC(=O)[C@H](CC(=O)N)N LJUOLNXOWSWGKF-ACZMJKKPSA-N 0.000 description 2
- NNMUHYLAYUSTTN-FXQIFTODSA-N Asn-Gln-Glu Chemical compound [H]N[C@@H](CC(N)=O)C(=O)N[C@@H](CCC(N)=O)C(=O)N[C@@H](CCC(O)=O)C(O)=O NNMUHYLAYUSTTN-FXQIFTODSA-N 0.000 description 2
- VJTWLBMESLDOMK-WDSKDSINSA-N Asn-Gln-Gly Chemical compound NC(=O)C[C@H](N)C(=O)N[C@@H](CCC(N)=O)C(=O)NCC(O)=O VJTWLBMESLDOMK-WDSKDSINSA-N 0.000 description 2
- GQRDIVQPSMPQME-ZPFDUUQYSA-N Asn-Ile-Leu Chemical compound [H]N[C@@H](CC(N)=O)C(=O)N[C@@H]([C@@H](C)CC)C(=O)N[C@@H](CC(C)C)C(O)=O GQRDIVQPSMPQME-ZPFDUUQYSA-N 0.000 description 2
- UHGUKCOQUNPSKK-CIUDSAMLSA-N Asn-Leu-Cys Chemical compound CC(C)C[C@@H](C(=O)N[C@@H](CS)C(=O)O)NC(=O)[C@H](CC(=O)N)N UHGUKCOQUNPSKK-CIUDSAMLSA-N 0.000 description 2
- XMHFCUKJRCQXGI-CIUDSAMLSA-N Asn-Pro-Gln Chemical compound C1C[C@H](N(C1)C(=O)[C@H](CC(=O)N)N)C(=O)N[C@@H](CCC(=O)N)C(=O)O XMHFCUKJRCQXGI-CIUDSAMLSA-N 0.000 description 2
- SZNGQSBRHFMZLT-IHRRRGAJSA-N Asn-Pro-Phe Chemical compound [H]N[C@@H](CC(N)=O)C(=O)N1CCC[C@H]1C(=O)N[C@@H](CC1=CC=CC=C1)C(O)=O SZNGQSBRHFMZLT-IHRRRGAJSA-N 0.000 description 2
- SUIJFTJDTJKSRK-IHRRRGAJSA-N Asn-Pro-Tyr Chemical compound NC(=O)C[C@H](N)C(=O)N1CCC[C@H]1C(=O)N[C@H](C(O)=O)CC1=CC=C(O)C=C1 SUIJFTJDTJKSRK-IHRRRGAJSA-N 0.000 description 2
- AMGQTNHANMRPOE-LKXGYXEUSA-N Asn-Thr-Ser Chemical compound [H]N[C@@H](CC(N)=O)C(=O)N[C@@H]([C@@H](C)O)C(=O)N[C@@H](CO)C(O)=O AMGQTNHANMRPOE-LKXGYXEUSA-N 0.000 description 2
- JNCRAQVYJZGIOW-QSFUFRPTSA-N Asn-Val-Ile Chemical compound [H]N[C@@H](CC(N)=O)C(=O)N[C@@H](C(C)C)C(=O)N[C@@H]([C@@H](C)CC)C(O)=O JNCRAQVYJZGIOW-QSFUFRPTSA-N 0.000 description 2
- XOQYDFCQPWAMSA-KKHAAJSZSA-N Asn-Val-Thr Chemical compound [H]N[C@@H](CC(N)=O)C(=O)N[C@@H](C(C)C)C(=O)N[C@@H]([C@@H](C)O)C(O)=O XOQYDFCQPWAMSA-KKHAAJSZSA-N 0.000 description 2
- FMWHSNJMHUNLAG-FXQIFTODSA-N Asp-Cys-Arg Chemical compound OC(=O)C[C@H](N)C(=O)N[C@@H](CS)C(=O)N[C@H](C(O)=O)CCCN=C(N)N FMWHSNJMHUNLAG-FXQIFTODSA-N 0.000 description 2
- XAJRHVUUVUPFQL-ACZMJKKPSA-N Asp-Glu-Asp Chemical compound OC(=O)C[C@H](N)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CC(O)=O)C(O)=O XAJRHVUUVUPFQL-ACZMJKKPSA-N 0.000 description 2
- GHODABZPVZMWCE-FXQIFTODSA-N Asp-Glu-Glu Chemical compound OC(=O)C[C@H](N)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CCC(O)=O)C(O)=O GHODABZPVZMWCE-FXQIFTODSA-N 0.000 description 2
- WWOYXVBGHAHQBG-FXQIFTODSA-N Asp-Met-Asp Chemical compound [H]N[C@@H](CC(O)=O)C(=O)N[C@@H](CCSC)C(=O)N[C@@H](CC(O)=O)C(O)=O WWOYXVBGHAHQBG-FXQIFTODSA-N 0.000 description 2
- BRRPVTUFESPTCP-ACZMJKKPSA-N Asp-Ser-Glu Chemical compound OC(=O)C[C@H](N)C(=O)N[C@@H](CO)C(=O)N[C@H](C(O)=O)CCC(O)=O BRRPVTUFESPTCP-ACZMJKKPSA-N 0.000 description 2
- YODBPLSWNJMZOJ-BPUTZDHNSA-N Asp-Trp-Arg Chemical compound C1=CC=C2C(=C1)C(=CN2)C[C@@H](C(=O)N[C@@H](CCCN=C(N)N)C(=O)O)NC(=O)[C@H](CC(=O)O)N YODBPLSWNJMZOJ-BPUTZDHNSA-N 0.000 description 2
- BOXNGMVEVOGXOJ-UBHSHLNASA-N Asp-Trp-Ser Chemical compound C1=CC=C2C(=C1)C(=CN2)C[C@@H](C(=O)N[C@@H](CO)C(=O)O)NC(=O)[C@H](CC(=O)O)N BOXNGMVEVOGXOJ-UBHSHLNASA-N 0.000 description 2
- XMKXONRMGJXCJV-LAEOZQHASA-N Asp-Val-Glu Chemical compound OC(=O)C[C@H](N)C(=O)N[C@@H](C(C)C)C(=O)N[C@@H](CCC(O)=O)C(O)=O XMKXONRMGJXCJV-LAEOZQHASA-N 0.000 description 2
- 208000035143 Bacterial infection Diseases 0.000 description 2
- 241000701489 Cauliflower mosaic virus Species 0.000 description 2
- 102000012286 Chitinases Human genes 0.000 description 2
- 108010022172 Chitinases Proteins 0.000 description 2
- 241001648414 Conospermum Species 0.000 description 2
- 235000019750 Crude protein Nutrition 0.000 description 2
- YMBAVNPKBWHDAW-CIUDSAMLSA-N Cys-Asp-Lys Chemical compound C(CCN)C[C@@H](C(=O)O)NC(=O)[C@H](CC(=O)O)NC(=O)[C@H](CS)N YMBAVNPKBWHDAW-CIUDSAMLSA-N 0.000 description 2
- ZVNFONSZVUBRAV-CIUDSAMLSA-N Cys-Gln-Arg Chemical compound C(C[C@@H](C(=O)O)NC(=O)[C@H](CCC(=O)N)NC(=O)[C@H](CS)N)CN=C(N)N ZVNFONSZVUBRAV-CIUDSAMLSA-N 0.000 description 2
- UDPSLLFHOLGXBY-FXQIFTODSA-N Cys-Glu-Glu Chemical compound [H]N[C@@H](CS)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CCC(O)=O)C(O)=O UDPSLLFHOLGXBY-FXQIFTODSA-N 0.000 description 2
- GUKYYUFHWYRMEU-WHFBIAKZSA-N Cys-Gly-Asp Chemical compound [H]N[C@@H](CS)C(=O)NCC(=O)N[C@@H](CC(O)=O)C(O)=O GUKYYUFHWYRMEU-WHFBIAKZSA-N 0.000 description 2
- MKMKILWCRQLDFJ-DCAQKATOSA-N Cys-Lys-Arg Chemical compound [H]N[C@@H](CS)C(=O)N[C@@H](CCCCN)C(=O)N[C@@H](CCCNC(N)=N)C(O)=O MKMKILWCRQLDFJ-DCAQKATOSA-N 0.000 description 2
- 108010002069 Defensins Proteins 0.000 description 2
- 102000000541 Defensins Human genes 0.000 description 2
- 208000035240 Disease Resistance Diseases 0.000 description 2
- YQYJSBFKSSDGFO-UHFFFAOYSA-N Epihygromycin Natural products OC1C(O)C(C(=O)C)OC1OC(C(=C1)O)=CC=C1C=C(C)C(=O)NC1C(O)C(O)C2OCOC2C1O YQYJSBFKSSDGFO-UHFFFAOYSA-N 0.000 description 2
- 108010074860 Factor Xa Proteins 0.000 description 2
- RGXXLQWXBFNXTG-CIUDSAMLSA-N Gln-Arg-Ala Chemical compound [H]N[C@@H](CCC(N)=O)C(=O)N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](C)C(O)=O RGXXLQWXBFNXTG-CIUDSAMLSA-N 0.000 description 2
- YNNXQZDEOCYJJL-CIUDSAMLSA-N Gln-Arg-Asp Chemical compound C(C[C@@H](C(=O)N[C@@H](CC(=O)O)C(=O)O)NC(=O)[C@H](CCC(=O)N)N)CN=C(N)N YNNXQZDEOCYJJL-CIUDSAMLSA-N 0.000 description 2
- WOACHWLUOFZLGJ-GUBZILKMSA-N Gln-Arg-Gln Chemical compound [H]N[C@@H](CCC(N)=O)C(=O)N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CCC(N)=O)C(O)=O WOACHWLUOFZLGJ-GUBZILKMSA-N 0.000 description 2
- XOKGKOQWADCLFQ-GARJFASQSA-N Gln-Arg-Pro Chemical compound C1C[C@@H](N(C1)C(=O)[C@H](CCCN=C(N)N)NC(=O)[C@H](CCC(=O)N)N)C(=O)O XOKGKOQWADCLFQ-GARJFASQSA-N 0.000 description 2
- RKAQZCDMSUQTSS-FXQIFTODSA-N Gln-Asp-Gln Chemical compound C(CC(=O)N)[C@@H](C(=O)N[C@@H](CC(=O)O)C(=O)N[C@@H](CCC(=O)N)C(=O)O)N RKAQZCDMSUQTSS-FXQIFTODSA-N 0.000 description 2
- IKDOHQHEFPPGJG-FXQIFTODSA-N Gln-Asp-Glu Chemical compound [H]N[C@@H](CCC(N)=O)C(=O)N[C@@H](CC(O)=O)C(=O)N[C@@H](CCC(O)=O)C(O)=O IKDOHQHEFPPGJG-FXQIFTODSA-N 0.000 description 2
- XFKUFUJECJUQTQ-CIUDSAMLSA-N Gln-Gln-Glu Chemical compound NC(=O)CC[C@H](N)C(=O)N[C@@H](CCC(N)=O)C(=O)N[C@@H](CCC(O)=O)C(O)=O XFKUFUJECJUQTQ-CIUDSAMLSA-N 0.000 description 2
- MADFVRSKEIEZHZ-DCAQKATOSA-N Gln-Gln-Lys Chemical compound C(CCN)C[C@@H](C(=O)O)NC(=O)[C@H](CCC(=O)N)NC(=O)[C@H](CCC(=O)N)N MADFVRSKEIEZHZ-DCAQKATOSA-N 0.000 description 2
- BLOXULLYFRGYKZ-GUBZILKMSA-N Gln-Glu-Arg Chemical compound [H]N[C@@H](CCC(N)=O)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CCCNC(N)=N)C(O)=O BLOXULLYFRGYKZ-GUBZILKMSA-N 0.000 description 2
- DDNIZQDYXDENIT-FXQIFTODSA-N Gln-Glu-Cys Chemical compound C(CC(=O)N)[C@@H](C(=O)N[C@@H](CCC(=O)O)C(=O)N[C@@H](CS)C(=O)O)N DDNIZQDYXDENIT-FXQIFTODSA-N 0.000 description 2
- LVSYIKGMLRHKME-IUCAKERBSA-N Gln-Gly-His Chemical compound C1=C(NC=N1)C[C@@H](C(=O)O)NC(=O)CNC(=O)[C@H](CCC(=O)N)N LVSYIKGMLRHKME-IUCAKERBSA-N 0.000 description 2
- SMLDOQHTOAAFJQ-WDSKDSINSA-N Gln-Gly-Ser Chemical compound [H]N[C@@H](CCC(N)=O)C(=O)NCC(=O)N[C@@H](CO)C(O)=O SMLDOQHTOAAFJQ-WDSKDSINSA-N 0.000 description 2
- HPCOBEHVEHWREJ-DCAQKATOSA-N Gln-Lys-Glu Chemical compound [H]N[C@@H](CCC(N)=O)C(=O)N[C@@H](CCCCN)C(=O)N[C@@H](CCC(O)=O)C(O)=O HPCOBEHVEHWREJ-DCAQKATOSA-N 0.000 description 2
- XUMFMAVDHQDATI-DCAQKATOSA-N Gln-Pro-Arg Chemical compound NC(=O)CC[C@H](N)C(=O)N1CCC[C@H]1C(=O)N[C@@H](CCCN=C(N)N)C(O)=O XUMFMAVDHQDATI-DCAQKATOSA-N 0.000 description 2
- CMBXOSFZCFGDLE-IHRRRGAJSA-N Gln-Tyr-Gln Chemical compound C1=CC(=CC=C1C[C@@H](C(=O)N[C@@H](CCC(=O)N)C(=O)O)NC(=O)[C@H](CCC(=O)N)N)O CMBXOSFZCFGDLE-IHRRRGAJSA-N 0.000 description 2
- FYBSCGZLICNOBA-XQXXSGGOSA-N Glu-Ala-Thr Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N[C@@H](C)C(=O)N[C@@H]([C@@H](C)O)C(O)=O FYBSCGZLICNOBA-XQXXSGGOSA-N 0.000 description 2
- WOMUDRVDJMHTCV-DCAQKATOSA-N Glu-Arg-Arg Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CCCNC(N)=N)C(O)=O WOMUDRVDJMHTCV-DCAQKATOSA-N 0.000 description 2
- DIXKFOPPGWKZLY-CIUDSAMLSA-N Glu-Arg-Asp Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CC(O)=O)C(O)=O DIXKFOPPGWKZLY-CIUDSAMLSA-N 0.000 description 2
- LXAUHIRMWXQRKI-XHNCKOQMSA-N Glu-Asn-Pro Chemical compound C1C[C@@H](N(C1)C(=O)[C@H](CC(=O)N)NC(=O)[C@H](CCC(=O)O)N)C(=O)O LXAUHIRMWXQRKI-XHNCKOQMSA-N 0.000 description 2
- NTBDVNJIWCKURJ-ACZMJKKPSA-N Glu-Asp-Asn Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N[C@@H](CC(O)=O)C(=O)N[C@@H](CC(N)=O)C(O)=O NTBDVNJIWCKURJ-ACZMJKKPSA-N 0.000 description 2
- NADWTMLCUDMDQI-ACZMJKKPSA-N Glu-Asp-Cys Chemical compound C(CC(=O)O)[C@@H](C(=O)N[C@@H](CC(=O)O)C(=O)N[C@@H](CS)C(=O)O)N NADWTMLCUDMDQI-ACZMJKKPSA-N 0.000 description 2
- IESFZVCAVACGPH-PEFMBERDSA-N Glu-Asp-Ile Chemical compound CC[C@H](C)[C@@H](C(O)=O)NC(=O)[C@H](CC(O)=O)NC(=O)[C@@H](N)CCC(O)=O IESFZVCAVACGPH-PEFMBERDSA-N 0.000 description 2
- WATXSTJXNBOHKD-LAEOZQHASA-N Glu-Asp-Val Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N[C@@H](CC(O)=O)C(=O)N[C@@H](C(C)C)C(O)=O WATXSTJXNBOHKD-LAEOZQHASA-N 0.000 description 2
- GFLQTABMFBXRIY-GUBZILKMSA-N Glu-Gln-Arg Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N[C@@H](CCC(N)=O)C(=O)N[C@@H](CCCNC(N)=N)C(O)=O GFLQTABMFBXRIY-GUBZILKMSA-N 0.000 description 2
- NKLRYVLERDYDBI-FXQIFTODSA-N Glu-Glu-Asp Chemical compound OC(=O)CC[C@H](N)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CC(O)=O)C(O)=O NKLRYVLERDYDBI-FXQIFTODSA-N 0.000 description 2
- IQACOVZVOMVILH-FXQIFTODSA-N Glu-Glu-Ser Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CO)C(O)=O IQACOVZVOMVILH-FXQIFTODSA-N 0.000 description 2
- PHONAZGUEGIOEM-GLLZPBPUSA-N Glu-Glu-Thr Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H]([C@@H](C)O)C(O)=O PHONAZGUEGIOEM-GLLZPBPUSA-N 0.000 description 2
- MTAOBYXRYJZRGQ-WDSKDSINSA-N Glu-Gly-Asp Chemical compound OC(=O)CC[C@H](N)C(=O)NCC(=O)N[C@@H](CC(O)=O)C(O)=O MTAOBYXRYJZRGQ-WDSKDSINSA-N 0.000 description 2
- OAGVHWYIBZMWLA-YFKPBYRVSA-N Glu-Gly-Gly Chemical compound OC(=O)CC[C@H](N)C(=O)NCC(=O)NCC(O)=O OAGVHWYIBZMWLA-YFKPBYRVSA-N 0.000 description 2
- XIKYNVKEUINBGL-IUCAKERBSA-N Glu-His-Gly Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N[C@@H](CC1=CNC=N1)C(=O)NCC(O)=O XIKYNVKEUINBGL-IUCAKERBSA-N 0.000 description 2
- ZHNHJYYFCGUZNQ-KBIXCLLPSA-N Glu-Ile-Ser Chemical compound OC[C@@H](C(O)=O)NC(=O)[C@H]([C@@H](C)CC)NC(=O)[C@@H](N)CCC(O)=O ZHNHJYYFCGUZNQ-KBIXCLLPSA-N 0.000 description 2
- HVYWQYLBVXMXSV-GUBZILKMSA-N Glu-Leu-Ala Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](C)C(O)=O HVYWQYLBVXMXSV-GUBZILKMSA-N 0.000 description 2
- UJMNFCAHLYKWOZ-DCAQKATOSA-N Glu-Lys-Gln Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N[C@@H](CCCCN)C(=O)N[C@@H](CCC(N)=O)C(O)=O UJMNFCAHLYKWOZ-DCAQKATOSA-N 0.000 description 2
- BCYGDJXHAGZNPQ-DCAQKATOSA-N Glu-Lys-Glu Chemical compound OC(=O)CC[C@H](N)C(=O)N[C@@H](CCCCN)C(=O)N[C@@H](CCC(O)=O)C(O)=O BCYGDJXHAGZNPQ-DCAQKATOSA-N 0.000 description 2
- ILWHFUZZCFYSKT-AVGNSLFASA-N Glu-Lys-Leu Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N[C@@H](CCCCN)C(=O)N[C@@H](CC(C)C)C(O)=O ILWHFUZZCFYSKT-AVGNSLFASA-N 0.000 description 2
- AAJHGGDRKHYSDH-GUBZILKMSA-N Glu-Pro-Gln Chemical compound C1C[C@H](N(C1)C(=O)[C@H](CCC(=O)O)N)C(=O)N[C@@H](CCC(=O)N)C(=O)O AAJHGGDRKHYSDH-GUBZILKMSA-N 0.000 description 2
- GMVCSRBOSIUTFC-FXQIFTODSA-N Glu-Ser-Glu Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N[C@@H](CO)C(=O)N[C@@H](CCC(O)=O)C(O)=O GMVCSRBOSIUTFC-FXQIFTODSA-N 0.000 description 2
- RFTVTKBHDXCEEX-WDSKDSINSA-N Glu-Ser-Gly Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N[C@@H](CO)C(=O)NCC(O)=O RFTVTKBHDXCEEX-WDSKDSINSA-N 0.000 description 2
- RXJFSLQVMGYQEL-IHRRRGAJSA-N Glu-Tyr-Glu Chemical compound OC(=O)CC[C@H](N)C(=O)N[C@H](C(=O)N[C@@H](CCC(O)=O)C(O)=O)CC1=CC=C(O)C=C1 RXJFSLQVMGYQEL-IHRRRGAJSA-N 0.000 description 2
- KFMBRBPXHVMDFN-UWVGGRQHSA-N Gly-Arg-Lys Chemical compound NCCCC[C@@H](C(O)=O)NC(=O)[C@@H](NC(=O)CN)CCCNC(N)=N KFMBRBPXHVMDFN-UWVGGRQHSA-N 0.000 description 2
- KQDMENMTYNBWMR-WHFBIAKZSA-N Gly-Asp-Ala Chemical compound [H]NCC(=O)N[C@@H](CC(O)=O)C(=O)N[C@@H](C)C(O)=O KQDMENMTYNBWMR-WHFBIAKZSA-N 0.000 description 2
- TZOVVRJYUDETQG-RCOVLWMOSA-N Gly-Asp-Val Chemical compound CC(C)[C@@H](C(O)=O)NC(=O)[C@H](CC(O)=O)NC(=O)CN TZOVVRJYUDETQG-RCOVLWMOSA-N 0.000 description 2
- XLFHCWHXKSFVIB-BQBZGAKWSA-N Gly-Gln-Gln Chemical compound NCC(=O)N[C@@H](CCC(N)=O)C(=O)N[C@@H](CCC(N)=O)C(O)=O XLFHCWHXKSFVIB-BQBZGAKWSA-N 0.000 description 2
- YFGONBOFGGWKKY-VHSXEESVSA-N Gly-His-Pro Chemical compound C1C[C@@H](N(C1)C(=O)[C@H](CC2=CN=CN2)NC(=O)CN)C(=O)O YFGONBOFGGWKKY-VHSXEESVSA-N 0.000 description 2
- IALQAMYQJBZNSK-WHFBIAKZSA-N Gly-Ser-Asn Chemical compound [H]NCC(=O)N[C@@H](CO)C(=O)N[C@@H](CC(N)=O)C(O)=O IALQAMYQJBZNSK-WHFBIAKZSA-N 0.000 description 2
- CSMYMGFCEJWALV-WDSKDSINSA-N Gly-Ser-Gln Chemical compound NCC(=O)N[C@@H](CO)C(=O)N[C@H](C(O)=O)CCC(N)=O CSMYMGFCEJWALV-WDSKDSINSA-N 0.000 description 2
- YABRDIBSPZONIY-BQBZGAKWSA-N Gly-Ser-Met Chemical compound [H]NCC(=O)N[C@@H](CO)C(=O)N[C@@H](CCSC)C(O)=O YABRDIBSPZONIY-BQBZGAKWSA-N 0.000 description 2
- ZVXMEWXHFBYJPI-LSJOCFKGSA-N Gly-Val-Ile Chemical compound [H]NCC(=O)N[C@@H](C(C)C)C(=O)N[C@@H]([C@@H](C)CC)C(O)=O ZVXMEWXHFBYJPI-LSJOCFKGSA-N 0.000 description 2
- 239000004471 Glycine Substances 0.000 description 2
- 244000299507 Gossypium hirsutum Species 0.000 description 2
- 235000009432 Gossypium hirsutum Nutrition 0.000 description 2
- KZTLOHBDLMIFSH-XVYDVKMFSA-N His-Ala-Asp Chemical compound [H]N[C@@H](CC1=CNC=N1)C(=O)N[C@@H](C)C(=O)N[C@@H](CC(O)=O)C(O)=O KZTLOHBDLMIFSH-XVYDVKMFSA-N 0.000 description 2
- FDQYIRHBVVUTJF-ZETCQYMHSA-N His-Gly-Gly Chemical compound [O-]C(=O)CNC(=O)CNC(=O)[C@@H]([NH3+])CC1=CN=CN1 FDQYIRHBVVUTJF-ZETCQYMHSA-N 0.000 description 2
- KYFGGRHWLFZXPU-KKUMJFAQSA-N His-Phe-Asn Chemical compound C1=CC=C(C=C1)C[C@@H](C(=O)N[C@@H](CC(=O)N)C(=O)O)NC(=O)[C@H](CC2=CN=CN2)N KYFGGRHWLFZXPU-KKUMJFAQSA-N 0.000 description 2
- ZHMZWSFQRUGLEC-JYJNAYRXSA-N His-Tyr-Glu Chemical compound [H]N[C@@H](CC1=CNC=N1)C(=O)N[C@@H](CC1=CC=C(O)C=C1)C(=O)N[C@@H](CCC(O)=O)C(O)=O ZHMZWSFQRUGLEC-JYJNAYRXSA-N 0.000 description 2
- TZCGZYWNIDZZMR-NAKRPEOUSA-N Ile-Arg-Ala Chemical compound CC[C@H](C)[C@@H](C(=O)N[C@@H](CCCN=C(N)N)C(=O)N[C@@H](C)C(=O)O)N TZCGZYWNIDZZMR-NAKRPEOUSA-N 0.000 description 2
- SCHZQZPYHBWYEQ-PEFMBERDSA-N Ile-Asn-Glu Chemical compound CC[C@H](C)[C@@H](C(=O)N[C@@H](CC(=O)N)C(=O)N[C@@H](CCC(=O)O)C(=O)O)N SCHZQZPYHBWYEQ-PEFMBERDSA-N 0.000 description 2
- OUUCIIJSBIBCHB-ZPFDUUQYSA-N Ile-Leu-Asp Chemical compound CC[C@H](C)[C@H](N)C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](CC(O)=O)C(O)=O OUUCIIJSBIBCHB-ZPFDUUQYSA-N 0.000 description 2
- XQLGNKLSPYCRMZ-HJWJTTGWSA-N Ile-Phe-Val Chemical compound CC[C@H](C)[C@@H](C(=O)N[C@@H](CC1=CC=CC=C1)C(=O)N[C@@H](C(C)C)C(=O)O)N XQLGNKLSPYCRMZ-HJWJTTGWSA-N 0.000 description 2
- IITVUURPOYGCTD-NAKRPEOUSA-N Ile-Pro-Ala Chemical compound CC[C@H](C)[C@H](N)C(=O)N1CCC[C@H]1C(=O)N[C@@H](C)C(O)=O IITVUURPOYGCTD-NAKRPEOUSA-N 0.000 description 2
- RQJUKVXWAKJDBW-SVSWQMSJSA-N Ile-Ser-Thr Chemical compound CC[C@H](C)[C@@H](C(=O)N[C@@H](CO)C(=O)N[C@@H]([C@@H](C)O)C(=O)O)N RQJUKVXWAKJDBW-SVSWQMSJSA-N 0.000 description 2
- DLEBSGAVWRPTIX-PEDHHIEDSA-N Ile-Val-Ile Chemical compound CC[C@H](C)[C@H](N)C(=O)N[C@@H](C(C)C)C(=O)N[C@H](C(O)=O)[C@@H](C)CC DLEBSGAVWRPTIX-PEDHHIEDSA-N 0.000 description 2
- IBMVEYRWAWIOTN-UHFFFAOYSA-N L-Leucyl-L-Arginyl-L-Proline Natural products CC(C)CC(N)C(=O)NC(CCCN=C(N)N)C(=O)N1CCCC1C(O)=O IBMVEYRWAWIOTN-UHFFFAOYSA-N 0.000 description 2
- QNAYBMKLOCPYGJ-REOHCLBHSA-N L-alanine Chemical compound C[C@H](N)C(O)=O QNAYBMKLOCPYGJ-REOHCLBHSA-N 0.000 description 2
- ODKSFYDXXFIFQN-BYPYZUCNSA-P L-argininium(2+) Chemical compound NC(=[NH2+])NCCC[C@H]([NH3+])C(O)=O ODKSFYDXXFIFQN-BYPYZUCNSA-P 0.000 description 2
- 241000228457 Leptosphaeria maculans Species 0.000 description 2
- CQQGCWPXDHTTNF-GUBZILKMSA-N Leu-Ala-Glu Chemical compound CC(C)C[C@H](N)C(=O)N[C@@H](C)C(=O)N[C@H](C(O)=O)CCC(O)=O CQQGCWPXDHTTNF-GUBZILKMSA-N 0.000 description 2
- XIRYQRLFHWWWTC-QEJZJMRPSA-N Leu-Ala-Phe Chemical compound CC(C)C[C@H](N)C(=O)N[C@@H](C)C(=O)N[C@H](C(O)=O)CC1=CC=CC=C1 XIRYQRLFHWWWTC-QEJZJMRPSA-N 0.000 description 2
- BQSLGJHIAGOZCD-CIUDSAMLSA-N Leu-Ala-Ser Chemical compound CC(C)C[C@H](N)C(=O)N[C@@H](C)C(=O)N[C@@H](CO)C(O)=O BQSLGJHIAGOZCD-CIUDSAMLSA-N 0.000 description 2
- REPPKAMYTOJTFC-DCAQKATOSA-N Leu-Arg-Asp Chemical compound [H]N[C@@H](CC(C)C)C(=O)N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CC(O)=O)C(O)=O REPPKAMYTOJTFC-DCAQKATOSA-N 0.000 description 2
- KSZCCRIGNVSHFH-UWVGGRQHSA-N Leu-Arg-Gly Chemical compound [H]N[C@@H](CC(C)C)C(=O)N[C@@H](CCCNC(N)=N)C(=O)NCC(O)=O KSZCCRIGNVSHFH-UWVGGRQHSA-N 0.000 description 2
- GPICTNQYKHHHTH-GUBZILKMSA-N Leu-Gln-Ser Chemical compound CC(C)C[C@H](N)C(=O)N[C@@H](CCC(N)=O)C(=O)N[C@@H](CO)C(O)=O GPICTNQYKHHHTH-GUBZILKMSA-N 0.000 description 2
- YVKSMSDXKMSIRX-GUBZILKMSA-N Leu-Glu-Asn Chemical compound [H]N[C@@H](CC(C)C)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CC(N)=O)C(O)=O YVKSMSDXKMSIRX-GUBZILKMSA-N 0.000 description 2
- FIYMBBHGYNQFOP-IUCAKERBSA-N Leu-Gly-Gln Chemical compound CC(C)C[C@@H](C(=O)NCC(=O)N[C@@H](CCC(=O)N)C(=O)O)N FIYMBBHGYNQFOP-IUCAKERBSA-N 0.000 description 2
- IFMPDNRWZZEZSL-SRVKXCTJSA-N Leu-Leu-Cys Chemical compound CC(C)C[C@H](N)C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](CS)C(O)=O IFMPDNRWZZEZSL-SRVKXCTJSA-N 0.000 description 2
- UBZGNBKMIJHOHL-BZSNNMDCSA-N Leu-Leu-Phe Chemical compound CC(C)C[C@H]([NH3+])C(=O)N[C@@H](CC(C)C)C(=O)N[C@H](C([O-])=O)CC1=CC=CC=C1 UBZGNBKMIJHOHL-BZSNNMDCSA-N 0.000 description 2
- ZDBMWELMUCLUPL-QEJZJMRPSA-N Leu-Phe-Ala Chemical compound CC(C)C[C@H](N)C(=O)N[C@H](C(=O)N[C@@H](C)C(O)=O)CC1=CC=CC=C1 ZDBMWELMUCLUPL-QEJZJMRPSA-N 0.000 description 2
- DRWMRVFCKKXHCH-BZSNNMDCSA-N Leu-Phe-Leu Chemical compound CC(C)C[C@H]([NH3+])C(=O)N[C@H](C(=O)N[C@@H](CC(C)C)C([O-])=O)CC1=CC=CC=C1 DRWMRVFCKKXHCH-BZSNNMDCSA-N 0.000 description 2
- PTRKPHUGYULXPU-KKUMJFAQSA-N Leu-Phe-Ser Chemical compound [H]N[C@@H](CC(C)C)C(=O)N[C@@H](CC1=CC=CC=C1)C(=O)N[C@@H](CO)C(O)=O PTRKPHUGYULXPU-KKUMJFAQSA-N 0.000 description 2
- DPURXCQCHSQPAN-AVGNSLFASA-N Leu-Pro-Pro Chemical compound CC(C)C[C@H](N)C(=O)N1CCC[C@H]1C(=O)N1[C@H](C(O)=O)CCC1 DPURXCQCHSQPAN-AVGNSLFASA-N 0.000 description 2
- SBANPBVRHYIMRR-GARJFASQSA-N Leu-Ser-Pro Chemical compound CC(C)C[C@@H](C(=O)N[C@@H](CO)C(=O)N1CCC[C@@H]1C(=O)O)N SBANPBVRHYIMRR-GARJFASQSA-N 0.000 description 2
- ODRREERHVHMIPT-OEAJRASXSA-N Leu-Thr-Phe Chemical compound CC(C)C[C@H](N)C(=O)N[C@@H]([C@@H](C)O)C(=O)N[C@H](C(O)=O)CC1=CC=CC=C1 ODRREERHVHMIPT-OEAJRASXSA-N 0.000 description 2
- VKVDRTGWLVZJOM-DCAQKATOSA-N Leu-Val-Ser Chemical compound CC(C)C[C@H](N)C(=O)N[C@@H](C(C)C)C(=O)N[C@@H](CO)C(O)=O VKVDRTGWLVZJOM-DCAQKATOSA-N 0.000 description 2
- QESXLSQLQHHTIX-RHYQMDGZSA-N Leu-Val-Thr Chemical compound CC(C)C[C@H](N)C(=O)N[C@@H](C(C)C)C(=O)N[C@@H]([C@@H](C)O)C(O)=O QESXLSQLQHHTIX-RHYQMDGZSA-N 0.000 description 2
- JGAMUXDWYSXYLM-SRVKXCTJSA-N Lys-Arg-Glu Chemical compound [H]N[C@@H](CCCCN)C(=O)N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CCC(O)=O)C(O)=O JGAMUXDWYSXYLM-SRVKXCTJSA-N 0.000 description 2
- BRSGXFITDXFMFF-IHRRRGAJSA-N Lys-Arg-His Chemical compound C1=C(NC=N1)C[C@@H](C(=O)O)NC(=O)[C@H](CCCN=C(N)N)NC(=O)[C@H](CCCCN)N BRSGXFITDXFMFF-IHRRRGAJSA-N 0.000 description 2
- GGAPIOORBXHMNY-ULQDDVLXSA-N Lys-Arg-Tyr Chemical compound C1=CC(=CC=C1C[C@@H](C(=O)O)NC(=O)[C@H](CCCN=C(N)N)NC(=O)[C@H](CCCCN)N)O GGAPIOORBXHMNY-ULQDDVLXSA-N 0.000 description 2
- NCTDKZKNBDZDOL-GARJFASQSA-N Lys-Asn-Pro Chemical compound C1C[C@@H](N(C1)C(=O)[C@H](CC(=O)N)NC(=O)[C@H](CCCCN)N)C(=O)O NCTDKZKNBDZDOL-GARJFASQSA-N 0.000 description 2
- MRWXLRGAFDOILG-DCAQKATOSA-N Lys-Gln-Gln Chemical compound [H]N[C@@H](CCCCN)C(=O)N[C@@H](CCC(N)=O)C(=O)N[C@@H](CCC(N)=O)C(O)=O MRWXLRGAFDOILG-DCAQKATOSA-N 0.000 description 2
- GRADYHMSAUIKPS-DCAQKATOSA-N Lys-Glu-Gln Chemical compound [H]N[C@@H](CCCCN)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CCC(N)=O)C(O)=O GRADYHMSAUIKPS-DCAQKATOSA-N 0.000 description 2
- GQZMPWBZQALKJO-UWVGGRQHSA-N Lys-Gly-Arg Chemical compound [H]N[C@@H](CCCCN)C(=O)NCC(=O)N[C@@H](CCCNC(N)=N)C(O)=O GQZMPWBZQALKJO-UWVGGRQHSA-N 0.000 description 2
- CANPXOLVTMKURR-WEDXCCLWSA-N Lys-Gly-Thr Chemical compound C[C@@H](O)[C@@H](C(O)=O)NC(=O)CNC(=O)[C@@H](N)CCCCN CANPXOLVTMKURR-WEDXCCLWSA-N 0.000 description 2
- WVJNGSFKBKOKRV-AJNGGQMLSA-N Lys-Leu-Ile Chemical compound [H]N[C@@H](CCCCN)C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H]([C@@H](C)CC)C(O)=O WVJNGSFKBKOKRV-AJNGGQMLSA-N 0.000 description 2
- RMOKGALPSPOYKE-KATARQTJSA-N Lys-Thr-Ser Chemical compound [H]N[C@@H](CCCCN)C(=O)N[C@@H]([C@@H](C)O)C(=O)N[C@@H](CO)C(O)=O RMOKGALPSPOYKE-KATARQTJSA-N 0.000 description 2
- RMKJOQSYLQQRFN-KKUMJFAQSA-N Lys-Tyr-Asp Chemical compound [H]N[C@@H](CCCCN)C(=O)N[C@@H](CC1=CC=C(O)C=C1)C(=O)N[C@@H](CC(O)=O)C(O)=O RMKJOQSYLQQRFN-KKUMJFAQSA-N 0.000 description 2
- CSNNHWWHGAXBCP-UHFFFAOYSA-L Magnesium sulfate Chemical compound [Mg+2].[O-][S+2]([O-])([O-])[O-] CSNNHWWHGAXBCP-UHFFFAOYSA-L 0.000 description 2
- 241000219071 Malvaceae Species 0.000 description 2
- QRHWTCJBCLGYRB-FXQIFTODSA-N Met-Ala-Cys Chemical compound CSCC[C@H](N)C(=O)N[C@@H](C)C(=O)N[C@@H](CS)C(O)=O QRHWTCJBCLGYRB-FXQIFTODSA-N 0.000 description 2
- QGQGAIBGTUJRBR-NAKRPEOUSA-N Met-Ala-Ile Chemical compound CC[C@H](C)[C@@H](C(O)=O)NC(=O)[C@H](C)NC(=O)[C@@H](N)CCSC QGQGAIBGTUJRBR-NAKRPEOUSA-N 0.000 description 2
- AWOMRHGUWFBDNU-ZPFDUUQYSA-N Met-Gln-Ile Chemical compound CC[C@H](C)[C@@H](C(=O)O)NC(=O)[C@H](CCC(=O)N)NC(=O)[C@H](CCSC)N AWOMRHGUWFBDNU-ZPFDUUQYSA-N 0.000 description 2
- MVBZBRKNZVJEKK-DTWKUNHWSA-N Met-Gly-Pro Chemical compound CSCC[C@@H](C(=O)NCC(=O)N1CCC[C@@H]1C(=O)O)N MVBZBRKNZVJEKK-DTWKUNHWSA-N 0.000 description 2
- WPTHAGXMYDRPFD-SRVKXCTJSA-N Met-Lys-Glu Chemical compound CSCC[C@H](N)C(=O)N[C@@H](CCCCN)C(=O)N[C@@H](CCC(O)=O)C(O)=O WPTHAGXMYDRPFD-SRVKXCTJSA-N 0.000 description 2
- KPVLLNDCBYXKNV-CYDGBPFRSA-N Met-Val-Ile Chemical compound [H]N[C@@H](CCSC)C(=O)N[C@@H](C(C)C)C(=O)N[C@@H]([C@@H](C)CC)C(O)=O KPVLLNDCBYXKNV-CYDGBPFRSA-N 0.000 description 2
- 239000012901 Milli-Q water Substances 0.000 description 2
- 208000031888 Mycoses Diseases 0.000 description 2
- XZFYRXDAULDNFX-UHFFFAOYSA-N N-L-cysteinyl-L-phenylalanine Natural products SCC(N)C(=O)NC(C(O)=O)CC1=CC=CC=C1 XZFYRXDAULDNFX-UHFFFAOYSA-N 0.000 description 2
- SEQKRHFRPICQDD-UHFFFAOYSA-N N-tris(hydroxymethyl)methylglycine Chemical compound OCC(CO)(CO)[NH2+]CC([O-])=O SEQKRHFRPICQDD-UHFFFAOYSA-N 0.000 description 2
- 101150008132 NDE1 gene Proteins 0.000 description 2
- 238000005481 NMR spectroscopy Methods 0.000 description 2
- 108020005187 Oligonucleotide Probes Proteins 0.000 description 2
- 108700026244 Open Reading Frames Proteins 0.000 description 2
- 101710084735 Osmotin-like protein Proteins 0.000 description 2
- ZWJKVFAYPLPCQB-UNQGMJICSA-N Phe-Arg-Thr Chemical compound C[C@@H](O)[C@H](NC(=O)[C@H](CCCN=C(N)N)NC(=O)[C@@H](N)Cc1ccccc1)C(O)=O ZWJKVFAYPLPCQB-UNQGMJICSA-N 0.000 description 2
- CDNPIRSCAFMMBE-SRVKXCTJSA-N Phe-Asn-Ser Chemical compound [H]N[C@@H](CC1=CC=CC=C1)C(=O)N[C@@H](CC(N)=O)C(=O)N[C@@H](CO)C(O)=O CDNPIRSCAFMMBE-SRVKXCTJSA-N 0.000 description 2
- LWPMGKSZPKFKJD-DZKIICNBSA-N Phe-Glu-Val Chemical compound [H]N[C@@H](CC1=CC=CC=C1)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](C(C)C)C(O)=O LWPMGKSZPKFKJD-DZKIICNBSA-N 0.000 description 2
- ZLGQEBCCANLYRA-RYUDHWBXSA-N Phe-Gly-Glu Chemical compound [H]N[C@@H](CC1=CC=CC=C1)C(=O)NCC(=O)N[C@@H](CCC(O)=O)C(O)=O ZLGQEBCCANLYRA-RYUDHWBXSA-N 0.000 description 2
- SPXWRYVHOZVYBU-ULQDDVLXSA-N Phe-His-Val Chemical compound CC(C)[C@@H](C(=O)O)NC(=O)[C@H](CC1=CN=CN1)NC(=O)[C@H](CC2=CC=CC=C2)N SPXWRYVHOZVYBU-ULQDDVLXSA-N 0.000 description 2
- YCCUXNNKXDGMAM-KKUMJFAQSA-N Phe-Leu-Ser Chemical compound [H]N[C@@H](CC1=CC=CC=C1)C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](CO)C(O)=O YCCUXNNKXDGMAM-KKUMJFAQSA-N 0.000 description 2
- OWSLLRKCHLTUND-BZSNNMDCSA-N Phe-Phe-Asn Chemical compound C1=CC=C(C=C1)C[C@@H](C(=O)N[C@@H](CC2=CC=CC=C2)C(=O)N[C@@H](CC(=O)N)C(=O)O)N OWSLLRKCHLTUND-BZSNNMDCSA-N 0.000 description 2
- GCFNFKNPCMBHNT-IRXDYDNUSA-N Phe-Tyr-Gly Chemical compound C1=CC=C(C=C1)C[C@@H](C(=O)N[C@@H](CC2=CC=C(C=C2)O)C(=O)NCC(=O)O)N GCFNFKNPCMBHNT-IRXDYDNUSA-N 0.000 description 2
- XALFIVXGQUEGKV-JSGCOSHPSA-N Phe-Val-Gly Chemical compound OC(=O)CNC(=O)[C@H](C(C)C)NC(=O)[C@@H](N)CC1=CC=CC=C1 XALFIVXGQUEGKV-JSGCOSHPSA-N 0.000 description 2
- JTKGCYOOJLUETJ-ULQDDVLXSA-N Phe-Val-Leu Chemical compound CC(C)C[C@@H](C(O)=O)NC(=O)[C@H](C(C)C)NC(=O)[C@@H](N)CC1=CC=CC=C1 JTKGCYOOJLUETJ-ULQDDVLXSA-N 0.000 description 2
- IEIFEYBAYFSRBQ-IHRRRGAJSA-N Phe-Val-Ser Chemical compound CC(C)[C@@H](C(=O)N[C@@H](CO)C(=O)O)NC(=O)[C@H](CC1=CC=CC=C1)N IEIFEYBAYFSRBQ-IHRRRGAJSA-N 0.000 description 2
- APZNYJFGVAGFCF-JYJNAYRXSA-N Phe-Val-Val Chemical compound CC(C)[C@H](NC(=O)[C@@H](NC(=O)[C@@H](N)Cc1ccccc1)C(C)C)C(O)=O APZNYJFGVAGFCF-JYJNAYRXSA-N 0.000 description 2
- HPXVFFIIGOAQRV-DCAQKATOSA-N Pro-Arg-Gln Chemical compound [H]N1CCC[C@H]1C(=O)N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CCC(N)=O)C(O)=O HPXVFFIIGOAQRV-DCAQKATOSA-N 0.000 description 2
- SWXSLPHTJVAWDF-VEVYYDQMSA-N Pro-Asn-Thr Chemical compound [H]N1CCC[C@H]1C(=O)N[C@@H](CC(N)=O)C(=O)N[C@@H]([C@@H](C)O)C(O)=O SWXSLPHTJVAWDF-VEVYYDQMSA-N 0.000 description 2
- VDGTVWFMRXVQCT-GUBZILKMSA-N Pro-Glu-Gln Chemical compound NC(=O)CC[C@@H](C(O)=O)NC(=O)[C@H](CCC(O)=O)NC(=O)[C@@H]1CCCN1 VDGTVWFMRXVQCT-GUBZILKMSA-N 0.000 description 2
- DMKWYMWNEKIPFC-IUCAKERBSA-N Pro-Gly-Arg Chemical compound [H]N1CCC[C@H]1C(=O)NCC(=O)N[C@@H](CCCNC(N)=N)C(O)=O DMKWYMWNEKIPFC-IUCAKERBSA-N 0.000 description 2
- STASJMBVVHNWCG-IHRRRGAJSA-N Pro-His-Leu Chemical compound C([C@@H](C(=O)N[C@@H](CC(C)C)C([O-])=O)NC(=O)[C@H]1[NH2+]CCC1)C1=CN=CN1 STASJMBVVHNWCG-IHRRRGAJSA-N 0.000 description 2
- XFFIGWGYMUFCCQ-ULQDDVLXSA-N Pro-His-Tyr Chemical compound C1=CC(O)=CC=C1C[C@@H](C([O-])=O)NC(=O)[C@@H](NC(=O)[C@H]1[NH2+]CCC1)CC1=CN=CN1 XFFIGWGYMUFCCQ-ULQDDVLXSA-N 0.000 description 2
- XYSXOCIWCPFOCG-IHRRRGAJSA-N Pro-Leu-Leu Chemical compound [H]N1CCC[C@H]1C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](CC(C)C)C(O)=O XYSXOCIWCPFOCG-IHRRRGAJSA-N 0.000 description 2
- AWQGDZBKQTYNMN-IHRRRGAJSA-N Pro-Phe-Asp Chemical compound C1C[C@H](NC1)C(=O)N[C@@H](CC2=CC=CC=C2)C(=O)N[C@@H](CC(=O)O)C(=O)O AWQGDZBKQTYNMN-IHRRRGAJSA-N 0.000 description 2
- ZAUHSLVPDLNTRZ-QXEWZRGKSA-N Pro-Val-Asn Chemical compound [H]N1CCC[C@H]1C(=O)N[C@@H](C(C)C)C(=O)N[C@@H](CC(N)=O)C(O)=O ZAUHSLVPDLNTRZ-QXEWZRGKSA-N 0.000 description 2
- ONIBWKKTOPOVIA-UHFFFAOYSA-N Proline Natural products OC(=O)C1CCCN1 ONIBWKKTOPOVIA-UHFFFAOYSA-N 0.000 description 2
- 108090000829 Ribosome Inactivating Proteins Proteins 0.000 description 2
- 240000004808 Saccharomyces cerevisiae Species 0.000 description 2
- 235000014680 Saccharomyces cerevisiae Nutrition 0.000 description 2
- 229920002684 Sepharose Polymers 0.000 description 2
- NLQUOHDCLSFABG-GUBZILKMSA-N Ser-Arg-Arg Chemical compound NC(N)=NCCC[C@H](NC(=O)[C@H](CO)N)C(=O)N[C@@H](CCCN=C(N)N)C(O)=O NLQUOHDCLSFABG-GUBZILKMSA-N 0.000 description 2
- FCRMLGJMPXCAHD-FXQIFTODSA-N Ser-Arg-Asn Chemical compound [H]N[C@@H](CO)C(=O)N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CC(N)=O)C(O)=O FCRMLGJMPXCAHD-FXQIFTODSA-N 0.000 description 2
- QEDMOZUJTGEIBF-FXQIFTODSA-N Ser-Arg-Asp Chemical compound [H]N[C@@H](CO)C(=O)N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CC(O)=O)C(O)=O QEDMOZUJTGEIBF-FXQIFTODSA-N 0.000 description 2
- VQBLHWSPVYYZTB-DCAQKATOSA-N Ser-Arg-His Chemical compound C1=C(NC=N1)C[C@@H](C(=O)O)NC(=O)[C@H](CCCN=C(N)N)NC(=O)[C@H](CO)N VQBLHWSPVYYZTB-DCAQKATOSA-N 0.000 description 2
- VGNYHOBZJKWRGI-CIUDSAMLSA-N Ser-Asn-Lys Chemical compound NCCCC[C@@H](C(O)=O)NC(=O)[C@H](CC(N)=O)NC(=O)[C@@H](N)CO VGNYHOBZJKWRGI-CIUDSAMLSA-N 0.000 description 2
- PVDTYLHUWAEYGY-CIUDSAMLSA-N Ser-Glu-Arg Chemical compound [H]N[C@@H](CO)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CCCNC(N)=N)C(O)=O PVDTYLHUWAEYGY-CIUDSAMLSA-N 0.000 description 2
- HJEBZBMOTCQYDN-ACZMJKKPSA-N Ser-Glu-Asp Chemical compound [H]N[C@@H](CO)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CC(O)=O)C(O)=O HJEBZBMOTCQYDN-ACZMJKKPSA-N 0.000 description 2
- DSGYZICNAMEJOC-AVGNSLFASA-N Ser-Glu-Phe Chemical compound [H]N[C@@H](CO)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CC1=CC=CC=C1)C(O)=O DSGYZICNAMEJOC-AVGNSLFASA-N 0.000 description 2
- SNVIOQXAHVORQM-WDSKDSINSA-N Ser-Gly-Gln Chemical compound [H]N[C@@H](CO)C(=O)NCC(=O)N[C@@H](CCC(N)=O)C(O)=O SNVIOQXAHVORQM-WDSKDSINSA-N 0.000 description 2
- VMLONWHIORGALA-SRVKXCTJSA-N Ser-Leu-Leu Chemical compound CC(C)C[C@@H](C([O-])=O)NC(=O)[C@H](CC(C)C)NC(=O)[C@@H]([NH3+])CO VMLONWHIORGALA-SRVKXCTJSA-N 0.000 description 2
- NUEHQDHDLDXCRU-GUBZILKMSA-N Ser-Pro-Arg Chemical compound OC[C@H](N)C(=O)N1CCC[C@H]1C(=O)N[C@@H](CCCN=C(N)N)C(O)=O NUEHQDHDLDXCRU-GUBZILKMSA-N 0.000 description 2
- DYEGLQRVMBWQLD-IXOXFDKPSA-N Ser-Thr-Phe Chemical compound C[C@H]([C@@H](C(=O)N[C@@H](CC1=CC=CC=C1)C(=O)O)NC(=O)[C@H](CO)N)O DYEGLQRVMBWQLD-IXOXFDKPSA-N 0.000 description 2
- SGZVZUCRAVSPKQ-FXQIFTODSA-N Ser-Val-Cys Chemical compound CC(C)[C@@H](C(=O)N[C@@H](CS)C(=O)O)NC(=O)[C@H](CO)N SGZVZUCRAVSPKQ-FXQIFTODSA-N 0.000 description 2
- ANOQEBQWIAYIMV-AEJSXWLSSA-N Ser-Val-Pro Chemical compound CC(C)[C@@H](C(=O)N1CCC[C@@H]1C(=O)O)NC(=O)[C@H](CO)N ANOQEBQWIAYIMV-AEJSXWLSSA-N 0.000 description 2
- MTCFGRXMJLQNBG-UHFFFAOYSA-N Serine Natural products OCC(N)C(O)=O MTCFGRXMJLQNBG-UHFFFAOYSA-N 0.000 description 2
- 241001648323 Stirlingia latifolia Species 0.000 description 2
- 108700005078 Synthetic Genes Proteins 0.000 description 2
- 241000721159 Thielaviopsis paradoxa Species 0.000 description 2
- GLQFKOVWXPPFTP-VEVYYDQMSA-N Thr-Arg-Asp Chemical compound [H]N[C@@H]([C@@H](C)O)C(=O)N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CC(O)=O)C(O)=O GLQFKOVWXPPFTP-VEVYYDQMSA-N 0.000 description 2
- JEDIEMIJYSRUBB-FOHZUACHSA-N Thr-Asp-Gly Chemical compound C[C@@H](O)[C@H](N)C(=O)N[C@@H](CC(O)=O)C(=O)NCC(O)=O JEDIEMIJYSRUBB-FOHZUACHSA-N 0.000 description 2
- UTCFSBBXPWKLTG-XKBZYTNZSA-N Thr-Cys-Gln Chemical compound C[C@H]([C@@H](C(=O)N[C@@H](CS)C(=O)N[C@@H](CCC(=O)N)C(=O)O)N)O UTCFSBBXPWKLTG-XKBZYTNZSA-N 0.000 description 2
- OYTNZCBFDXGQGE-XQXXSGGOSA-N Thr-Gln-Ala Chemical compound C[C@H]([C@@H](C(=O)N[C@@H](CCC(=O)N)C(=O)N[C@@H](C)C(=O)O)N)O OYTNZCBFDXGQGE-XQXXSGGOSA-N 0.000 description 2
- KGKWKSSSQGGYAU-SUSMZKCASA-N Thr-Gln-Thr Chemical compound C[C@H]([C@@H](C(=O)N[C@@H](CCC(=O)N)C(=O)N[C@@H]([C@@H](C)O)C(=O)O)N)O KGKWKSSSQGGYAU-SUSMZKCASA-N 0.000 description 2
- AYCQVUUPIJHJTA-IXOXFDKPSA-N Thr-His-Leu Chemical compound [H]N[C@@H]([C@@H](C)O)C(=O)N[C@@H](CC1=CNC=N1)C(=O)N[C@@H](CC(C)C)C(O)=O AYCQVUUPIJHJTA-IXOXFDKPSA-N 0.000 description 2
- FQPDRTDDEZXCEC-SVSWQMSJSA-N Thr-Ile-Ser Chemical compound [H]N[C@@H]([C@@H](C)O)C(=O)N[C@@H]([C@@H](C)CC)C(=O)N[C@@H](CO)C(O)=O FQPDRTDDEZXCEC-SVSWQMSJSA-N 0.000 description 2
- MXDOAJQRJBMGMO-FJXKBIBVSA-N Thr-Pro-Gly Chemical compound C[C@@H](O)[C@H](N)C(=O)N1CCC[C@H]1C(=O)NCC(O)=O MXDOAJQRJBMGMO-FJXKBIBVSA-N 0.000 description 2
- MNYNCKZAEIAONY-XGEHTFHBSA-N Thr-Val-Ser Chemical compound C[C@@H](O)[C@H](N)C(=O)N[C@@H](C(C)C)C(=O)N[C@@H](CO)C(O)=O MNYNCKZAEIAONY-XGEHTFHBSA-N 0.000 description 2
- CKKFTIQYURNSEI-IHRRRGAJSA-N Tyr-Asn-Arg Chemical compound NC(N)=NCCC[C@@H](C(O)=O)NC(=O)[C@H](CC(N)=O)NC(=O)[C@@H](N)CC1=CC=C(O)C=C1 CKKFTIQYURNSEI-IHRRRGAJSA-N 0.000 description 2
- MNMYOSZWCKYEDI-JRQIVUDYSA-N Tyr-Asp-Thr Chemical compound [H]N[C@@H](CC1=CC=C(O)C=C1)C(=O)N[C@@H](CC(O)=O)C(=O)N[C@@H]([C@@H](C)O)C(O)=O MNMYOSZWCKYEDI-JRQIVUDYSA-N 0.000 description 2
- JWGXUKHIKXZWNG-RYUDHWBXSA-N Tyr-Gly-Gln Chemical compound C1=CC(=CC=C1C[C@@H](C(=O)NCC(=O)N[C@@H](CCC(=O)N)C(=O)O)N)O JWGXUKHIKXZWNG-RYUDHWBXSA-N 0.000 description 2
- LMKKMCGTDANZTR-BZSNNMDCSA-N Tyr-Phe-Asp Chemical compound C([C@H](N)C(=O)N[C@@H](CC=1C=CC=CC=1)C(=O)N[C@@H](CC(O)=O)C(O)=O)C1=CC=C(O)C=C1 LMKKMCGTDANZTR-BZSNNMDCSA-N 0.000 description 2
- WURLIFOWSMBUAR-SLFFLAALSA-N Tyr-Phe-Pro Chemical compound C1C[C@@H](N(C1)C(=O)[C@H](CC2=CC=CC=C2)NC(=O)[C@H](CC3=CC=C(C=C3)O)N)C(=O)O WURLIFOWSMBUAR-SLFFLAALSA-N 0.000 description 2
- NVJCMGGZHOJNBU-UFYCRDLUSA-N Tyr-Val-Phe Chemical compound CC(C)[C@@H](C(=O)N[C@@H](CC1=CC=CC=C1)C(=O)O)NC(=O)[C@H](CC2=CC=C(C=C2)O)N NVJCMGGZHOJNBU-UFYCRDLUSA-N 0.000 description 2
- UEOOXDLMQZBPFR-ZKWXMUAHSA-N Val-Ala-Asn Chemical compound C[C@@H](C(=O)N[C@@H](CC(=O)N)C(=O)O)NC(=O)[C@H](C(C)C)N UEOOXDLMQZBPFR-ZKWXMUAHSA-N 0.000 description 2
- YFOCMOVJBQDBCE-NRPADANISA-N Val-Ala-Glu Chemical compound C[C@@H](C(=O)N[C@@H](CCC(=O)O)C(=O)O)NC(=O)[C@H](C(C)C)N YFOCMOVJBQDBCE-NRPADANISA-N 0.000 description 2
- ASQFIHTXXMFENG-XPUUQOCRSA-N Val-Ala-Gly Chemical compound CC(C)[C@H](N)C(=O)N[C@@H](C)C(=O)NCC(O)=O ASQFIHTXXMFENG-XPUUQOCRSA-N 0.000 description 2
- RUCNAYOMFXRIKJ-DCAQKATOSA-N Val-Ala-Lys Chemical compound CC(C)[C@H](N)C(=O)N[C@@H](C)C(=O)N[C@H](C(O)=O)CCCCN RUCNAYOMFXRIKJ-DCAQKATOSA-N 0.000 description 2
- ZLFHAAGHGQBQQN-GUBZILKMSA-N Val-Ala-Pro Natural products CC(C)[C@H](N)C(=O)N[C@@H](C)C(=O)N1CCC[C@H]1C(O)=O ZLFHAAGHGQBQQN-GUBZILKMSA-N 0.000 description 2
- VLOYGOZDPGYWFO-LAEOZQHASA-N Val-Asp-Glu Chemical compound CC(C)[C@H](N)C(=O)N[C@@H](CC(O)=O)C(=O)N[C@@H](CCC(O)=O)C(O)=O VLOYGOZDPGYWFO-LAEOZQHASA-N 0.000 description 2
- XTAUQCGQFJQGEJ-NHCYSSNCSA-N Val-Gln-Arg Chemical compound CC(C)[C@@H](C(=O)N[C@@H](CCC(=O)N)C(=O)N[C@@H](CCCN=C(N)N)C(=O)O)N XTAUQCGQFJQGEJ-NHCYSSNCSA-N 0.000 description 2
- QHFQQRKNGCXTHL-AUTRQRHGSA-N Val-Gln-Glu Chemical compound CC(C)[C@H](N)C(=O)N[C@@H](CCC(N)=O)C(=O)N[C@@H](CCC(O)=O)C(O)=O QHFQQRKNGCXTHL-AUTRQRHGSA-N 0.000 description 2
- CPTQYHDSVGVGDZ-UKJIMTQDSA-N Val-Gln-Ile Chemical compound CC[C@H](C)[C@@H](C(=O)O)NC(=O)[C@H](CCC(=O)N)NC(=O)[C@H](C(C)C)N CPTQYHDSVGVGDZ-UKJIMTQDSA-N 0.000 description 2
- KDKLLPMFFGYQJD-CYDGBPFRSA-N Val-Ile-Arg Chemical compound CC[C@H](C)[C@@H](C(=O)N[C@@H](CCCN=C(N)N)C(=O)O)NC(=O)[C@H](C(C)C)N KDKLLPMFFGYQJD-CYDGBPFRSA-N 0.000 description 2
- UKEVLVBHRKWECS-LSJOCFKGSA-N Val-Ile-Gly Chemical compound CC[C@H](C)[C@@H](C(=O)NCC(=O)O)NC(=O)[C@H](C(C)C)N UKEVLVBHRKWECS-LSJOCFKGSA-N 0.000 description 2
- VHRLUTIMTDOVCG-PEDHHIEDSA-N Val-Ile-Ile Chemical compound CC[C@H](C)[C@@H](C(=O)N[C@@H]([C@@H](C)CC)C(=O)O)NC(=O)[C@H](C(C)C)N VHRLUTIMTDOVCG-PEDHHIEDSA-N 0.000 description 2
- OTJMMKPMLUNTQT-AVGNSLFASA-N Val-Leu-Arg Chemical compound CC(C)C[C@@H](C(=O)N[C@@H](CCCN=C(N)N)C(=O)O)NC(=O)[C@H](C(C)C)N OTJMMKPMLUNTQT-AVGNSLFASA-N 0.000 description 2
- LYERIXUFCYVFFX-GVXVVHGQSA-N Val-Leu-Glu Chemical compound CC(C)C[C@@H](C(=O)N[C@@H](CCC(=O)O)C(=O)O)NC(=O)[C@H](C(C)C)N LYERIXUFCYVFFX-GVXVVHGQSA-N 0.000 description 2
- IJGPOONOTBNTFS-GVXVVHGQSA-N Val-Lys-Glu Chemical compound [H]N[C@@H](C(C)C)C(=O)N[C@@H](CCCCN)C(=O)N[C@@H](CCC(O)=O)C(O)=O IJGPOONOTBNTFS-GVXVVHGQSA-N 0.000 description 2
- XBJKAZATRJBDCU-GUBZILKMSA-N Val-Pro-Ala Chemical compound CC(C)[C@H](N)C(=O)N1CCC[C@H]1C(=O)N[C@@H](C)C(O)=O XBJKAZATRJBDCU-GUBZILKMSA-N 0.000 description 2
- DEGUERSKQBRZMZ-FXQIFTODSA-N Val-Ser-Ala Chemical compound CC(C)[C@H](N)C(=O)N[C@@H](CO)C(=O)N[C@@H](C)C(O)=O DEGUERSKQBRZMZ-FXQIFTODSA-N 0.000 description 2
- RYHUIHUOYRNNIE-NRPADANISA-N Val-Ser-Gln Chemical compound CC(C)[C@@H](C(=O)N[C@@H](CO)C(=O)N[C@@H](CCC(=O)N)C(=O)O)N RYHUIHUOYRNNIE-NRPADANISA-N 0.000 description 2
- GBIUHAYJGWVNLN-UHFFFAOYSA-N Val-Ser-Pro Natural products CC(C)C(N)C(=O)NC(CO)C(=O)N1CCCC1C(O)=O GBIUHAYJGWVNLN-UHFFFAOYSA-N 0.000 description 2
- PZTZYZUTCPZWJH-FXQIFTODSA-N Val-Ser-Ser Chemical compound CC(C)[C@@H](C(=O)N[C@@H](CO)C(=O)N[C@@H](CO)C(=O)O)N PZTZYZUTCPZWJH-FXQIFTODSA-N 0.000 description 2
- DFQZDQPLWBSFEJ-LSJOCFKGSA-N Val-Val-Asn Chemical compound CC(C)[C@@H](C(=O)N[C@@H](C(C)C)C(=O)N[C@@H](CC(=O)N)C(=O)O)N DFQZDQPLWBSFEJ-LSJOCFKGSA-N 0.000 description 2
- WBPFYNYTYASCQP-CYDGBPFRSA-N Val-Val-Ile Chemical compound CC[C@H](C)[C@@H](C(=O)O)NC(=O)[C@H](C(C)C)NC(=O)[C@H](C(C)C)N WBPFYNYTYASCQP-CYDGBPFRSA-N 0.000 description 2
- XNLUVJPMPAZHCY-JYJNAYRXSA-N Val-Val-Phe Chemical compound CC(C)[C@H]([NH3+])C(=O)N[C@@H](C(C)C)C(=O)N[C@H](C([O-])=O)CC1=CC=CC=C1 XNLUVJPMPAZHCY-JYJNAYRXSA-N 0.000 description 2
- 235000004279 alanine Nutrition 0.000 description 2
- 108010011559 alanylphenylalanine Proteins 0.000 description 2
- 238000011483 antifungal activity assay Methods 0.000 description 2
- ODKSFYDXXFIFQN-UHFFFAOYSA-N arginine Natural products OC(=O)C(N)CCCNC(N)=N ODKSFYDXXFIFQN-UHFFFAOYSA-N 0.000 description 2
- 108010069926 arginyl-glycyl-serine Proteins 0.000 description 2
- 238000003556 assay Methods 0.000 description 2
- 230000001580 bacterial effect Effects 0.000 description 2
- 210000004899 c-terminal region Anatomy 0.000 description 2
- 230000015556 catabolic process Effects 0.000 description 2
- 238000005277 cation exchange chromatography Methods 0.000 description 2
- 238000004113 cell culture Methods 0.000 description 2
- 239000006285 cell suspension Substances 0.000 description 2
- 108050003126 conotoxin Proteins 0.000 description 2
- 238000006731 degradation reaction Methods 0.000 description 2
- 238000013461 design Methods 0.000 description 2
- 238000000502 dialysis Methods 0.000 description 2
- FSXRLASFHBWESK-UHFFFAOYSA-N dipeptide phenylalanyl-tyrosine Natural products C=1C=C(O)C=CC=1CC(C(O)=O)NC(=O)C(N)CC1=CC=CC=C1 FSXRLASFHBWESK-UHFFFAOYSA-N 0.000 description 2
- 238000004520 electroporation Methods 0.000 description 2
- 238000002474 experimental method Methods 0.000 description 2
- 239000011536 extraction buffer Substances 0.000 description 2
- 235000013305 food Nutrition 0.000 description 2
- 108020001507 fusion proteins Proteins 0.000 description 2
- 102000037865 fusion proteins Human genes 0.000 description 2
- 108010006664 gamma-glutamyl-glycyl-glycine Proteins 0.000 description 2
- 108010079547 glutamylmethionine Proteins 0.000 description 2
- JYPCXBJRLBHWME-UHFFFAOYSA-N glycyl-L-prolyl-L-arginine Natural products NCC(=O)N1CCCC1C(=O)NC(CCCN=C(N)N)C(O)=O JYPCXBJRLBHWME-UHFFFAOYSA-N 0.000 description 2
- 108010000434 glycyl-alanyl-leucine Proteins 0.000 description 2
- 108010015792 glycyllysine Proteins 0.000 description 2
- 108010081551 glycylphenylalanine Proteins 0.000 description 2
- 230000009036 growth inhibition Effects 0.000 description 2
- HNDVDQJCIGZPNO-UHFFFAOYSA-N histidine Natural products OC(=O)C(N)CC1=CN=CN1 HNDVDQJCIGZPNO-UHFFFAOYSA-N 0.000 description 2
- 108010085325 histidylproline Proteins 0.000 description 2
- 238000000338 in vitro Methods 0.000 description 2
- 230000001939 inductive effect Effects 0.000 description 2
- 238000002955 isolation Methods 0.000 description 2
- 108010076756 leucyl-alanyl-phenylalanine Proteins 0.000 description 2
- 108010073472 leucyl-prolyl-proline Proteins 0.000 description 2
- 108010057821 leucylproline Proteins 0.000 description 2
- 239000007788 liquid Substances 0.000 description 2
- 239000003550 marker Substances 0.000 description 2
- 238000004949 mass spectrometry Methods 0.000 description 2
- 239000002609 medium Substances 0.000 description 2
- 238000002156 mixing Methods 0.000 description 2
- 210000004897 n-terminal region Anatomy 0.000 description 2
- 235000014571 nuts Nutrition 0.000 description 2
- 239000002751 oligonucleotide probe Substances 0.000 description 2
- 108010089198 phenylalanyl-prolyl-arginine Proteins 0.000 description 2
- 108010084572 phenylalanyl-valine Proteins 0.000 description 2
- 108010012581 phenylalanylglutamate Proteins 0.000 description 2
- 230000008488 polyadenylation Effects 0.000 description 2
- 230000003389 potentiating effect Effects 0.000 description 2
- 239000002243 precursor Substances 0.000 description 2
- 230000008569 process Effects 0.000 description 2
- 108010031719 prolyl-serine Proteins 0.000 description 2
- 108010090894 prolylleucine Proteins 0.000 description 2
- 238000004366 reverse phase liquid chromatography Methods 0.000 description 2
- 210000002966 serum Anatomy 0.000 description 2
- 108010048397 seryl-lysyl-leucine Proteins 0.000 description 2
- 239000011780 sodium chloride Substances 0.000 description 2
- 229910000162 sodium phosphate Inorganic materials 0.000 description 2
- 229940074404 sodium succinate Drugs 0.000 description 2
- ZDQYSKICYIVCPN-UHFFFAOYSA-L sodium succinate (anhydrous) Chemical compound [Na+].[Na+].[O-]C(=O)CCC([O-])=O ZDQYSKICYIVCPN-UHFFFAOYSA-L 0.000 description 2
- 239000002195 soluble material Substances 0.000 description 2
- 239000002904 solvent Substances 0.000 description 2
- 239000007858 starting material Substances 0.000 description 2
- 239000012085 test solution Substances 0.000 description 2
- 238000012546 transfer Methods 0.000 description 2
- WFKWXMTUELFFGS-UHFFFAOYSA-N tungsten Chemical compound [W] WFKWXMTUELFFGS-UHFFFAOYSA-N 0.000 description 2
- 229910052721 tungsten Inorganic materials 0.000 description 2
- 239000010937 tungsten Substances 0.000 description 2
- 108010073969 valyllysine Proteins 0.000 description 2
- XLYOFNOQVPJJNP-UHFFFAOYSA-N water Substances O XLYOFNOQVPJJNP-UHFFFAOYSA-N 0.000 description 2
- MTCFGRXMJLQNBG-REOHCLBHSA-N (2S)-2-Amino-3-hydroxypropansäure Chemical compound OC[C@H](N)C(O)=O MTCFGRXMJLQNBG-REOHCLBHSA-N 0.000 description 1
- NTUPOKHATNSWCY-PMPSAXMXSA-N (2s)-2-[[(2s)-1-[(2r)-2-amino-3-phenylpropanoyl]pyrrolidine-2-carbonyl]amino]-5-(diaminomethylideneamino)pentanoic acid Chemical compound C([C@@H](N)C(=O)N1[C@@H](CCC1)C(=O)N[C@@H](CCCN=C(N)N)C(O)=O)C1=CC=CC=C1 NTUPOKHATNSWCY-PMPSAXMXSA-N 0.000 description 1
- QRXMUCSWCMTJGU-UHFFFAOYSA-L (5-bromo-4-chloro-1h-indol-3-yl) phosphate Chemical compound C1=C(Br)C(Cl)=C2C(OP([O-])(=O)[O-])=CNC2=C1 QRXMUCSWCMTJGU-UHFFFAOYSA-L 0.000 description 1
- PIDRBUDUWHBYSR-UHFFFAOYSA-N 1-[2-[[2-[(2-amino-4-methylpentanoyl)amino]-4-methylpentanoyl]amino]-4-methylpentanoyl]pyrrolidine-2-carboxylic acid Chemical compound CC(C)CC(N)C(=O)NC(CC(C)C)C(=O)NC(CC(C)C)C(=O)N1CCCC1C(O)=O PIDRBUDUWHBYSR-UHFFFAOYSA-N 0.000 description 1
- YEJQWBFDKKTPNO-UHFFFAOYSA-N 2-[[2-[[1-(2-amino-3-methylbutanoyl)pyrrolidine-2-carbonyl]amino]acetyl]amino]-3-methylbutanoic acid Chemical compound CC(C)C(N)C(=O)N1CCCC1C(=O)NCC(=O)NC(C(C)C)C(O)=O YEJQWBFDKKTPNO-UHFFFAOYSA-N 0.000 description 1
- JUEUYDRZJNQZGR-UHFFFAOYSA-N 2-[[2-[[2-[(2-amino-4-methylpentanoyl)amino]-4-methylpentanoyl]amino]acetyl]amino]-3-phenylpropanoic acid Chemical compound CC(C)CC(N)C(=O)NC(CC(C)C)C(=O)NCC(=O)NC(C(O)=O)CC1=CC=CC=C1 JUEUYDRZJNQZGR-UHFFFAOYSA-N 0.000 description 1
- FMYBFLOWKQRBST-UHFFFAOYSA-N 2-[bis(carboxymethyl)amino]acetic acid;nickel Chemical compound [Ni].OC(=O)CN(CC(O)=O)CC(O)=O FMYBFLOWKQRBST-UHFFFAOYSA-N 0.000 description 1
- QFVHZQCOUORWEI-UHFFFAOYSA-N 4-[(4-anilino-5-sulfonaphthalen-1-yl)diazenyl]-5-hydroxynaphthalene-2,7-disulfonic acid Chemical compound C=12C(O)=CC(S(O)(=O)=O)=CC2=CC(S(O)(=O)=O)=CC=1N=NC(C1=CC=CC(=C11)S(O)(=O)=O)=CC=C1NC1=CC=CC=C1 QFVHZQCOUORWEI-UHFFFAOYSA-N 0.000 description 1
- 125000004042 4-aminobutyl group Chemical group [H]C([*])([H])C([H])([H])C([H])([H])C([H])([H])N([H])[H] 0.000 description 1
- IMIZPWSVYADSCN-UHFFFAOYSA-N 4-methyl-2-[[4-methyl-2-[[4-methyl-2-(pyrrolidine-2-carbonylamino)pentanoyl]amino]pentanoyl]amino]pentanoic acid Chemical compound CC(C)CC(C(O)=O)NC(=O)C(CC(C)C)NC(=O)C(CC(C)C)NC(=O)C1CCCN1 IMIZPWSVYADSCN-UHFFFAOYSA-N 0.000 description 1
- 101150073246 AGL1 gene Proteins 0.000 description 1
- QTBSBXVTEAMEQO-UHFFFAOYSA-M Acetate Chemical compound CC([O-])=O QTBSBXVTEAMEQO-UHFFFAOYSA-M 0.000 description 1
- 108010001604 Aesculus hippocastanum Ah-AMP1protein Proteins 0.000 description 1
- 229920000936 Agarose Polymers 0.000 description 1
- 241000589156 Agrobacterium rhizogenes Species 0.000 description 1
- PIPTUBPKYFRLCP-NHCYSSNCSA-N Ala-Ala-Phe Chemical compound C[C@H](N)C(=O)N[C@@H](C)C(=O)N[C@H](C(O)=O)CC1=CC=CC=C1 PIPTUBPKYFRLCP-NHCYSSNCSA-N 0.000 description 1
- YYSWCHMLFJLLBJ-ZLUOBGJFSA-N Ala-Ala-Ser Chemical compound C[C@H](N)C(=O)N[C@@H](C)C(=O)N[C@@H](CO)C(O)=O YYSWCHMLFJLLBJ-ZLUOBGJFSA-N 0.000 description 1
- WRDANSJTFOHBPI-FXQIFTODSA-N Ala-Arg-Cys Chemical compound C[C@@H](C(=O)N[C@@H](CCCN=C(N)N)C(=O)N[C@@H](CS)C(=O)O)N WRDANSJTFOHBPI-FXQIFTODSA-N 0.000 description 1
- SKHCUBQVZJHOFM-NAKRPEOUSA-N Ala-Arg-Ile Chemical compound [H]N[C@@H](C)C(=O)N[C@@H](CCCNC(N)=N)C(=O)N[C@@H]([C@@H](C)CC)C(O)=O SKHCUBQVZJHOFM-NAKRPEOUSA-N 0.000 description 1
- IMMKUCQIKKXKNP-DCAQKATOSA-N Ala-Arg-Leu Chemical compound CC(C)C[C@@H](C(O)=O)NC(=O)[C@@H](NC(=O)[C@H](C)N)CCCN=C(N)N IMMKUCQIKKXKNP-DCAQKATOSA-N 0.000 description 1
- YWWATNIVMOCSAV-UBHSHLNASA-N Ala-Arg-Phe Chemical compound NC(=N)NCCC[C@H](NC(=O)[C@@H](N)C)C(=O)N[C@H](C(O)=O)CC1=CC=CC=C1 YWWATNIVMOCSAV-UBHSHLNASA-N 0.000 description 1
- UCIYCBSJBQGDGM-LPEHRKFASA-N Ala-Arg-Pro Chemical compound C[C@@H](C(=O)N[C@@H](CCCN=C(N)N)C(=O)N1CCC[C@@H]1C(=O)O)N UCIYCBSJBQGDGM-LPEHRKFASA-N 0.000 description 1
- KUDREHRZRIVKHS-UWJYBYFXSA-N Ala-Asp-Tyr Chemical compound [H]N[C@@H](C)C(=O)N[C@@H](CC(O)=O)C(=O)N[C@@H](CC1=CC=C(O)C=C1)C(O)=O KUDREHRZRIVKHS-UWJYBYFXSA-N 0.000 description 1
- WCBVQNZTOKJWJS-ACZMJKKPSA-N Ala-Cys-Glu Chemical compound [H]N[C@@H](C)C(=O)N[C@@H](CS)C(=O)N[C@@H](CCC(O)=O)C(O)=O WCBVQNZTOKJWJS-ACZMJKKPSA-N 0.000 description 1
- VIGKUFXFTPWYER-BIIVOSGPSA-N Ala-Cys-Pro Chemical compound C[C@@H](C(=O)N[C@@H](CS)C(=O)N1CCC[C@@H]1C(=O)O)N VIGKUFXFTPWYER-BIIVOSGPSA-N 0.000 description 1
- CSAHOYQKNHGDHX-ACZMJKKPSA-N Ala-Gln-Asn Chemical compound C[C@H](N)C(=O)N[C@@H](CCC(N)=O)C(=O)N[C@@H](CC(N)=O)C(O)=O CSAHOYQKNHGDHX-ACZMJKKPSA-N 0.000 description 1
- RXTBLQVXNIECFP-FXQIFTODSA-N Ala-Gln-Gln Chemical compound C[C@H](N)C(=O)N[C@@H](CCC(N)=O)C(=O)N[C@@H](CCC(N)=O)C(O)=O RXTBLQVXNIECFP-FXQIFTODSA-N 0.000 description 1
- ZODMADSIQZZBSQ-FXQIFTODSA-N Ala-Gln-Glu Chemical compound [H]N[C@@H](C)C(=O)N[C@@H](CCC(N)=O)C(=O)N[C@@H](CCC(O)=O)C(O)=O ZODMADSIQZZBSQ-FXQIFTODSA-N 0.000 description 1
- IFTVANMRTIHKML-WDSKDSINSA-N Ala-Gln-Gly Chemical compound C[C@H](N)C(=O)N[C@@H](CCC(N)=O)C(=O)NCC(O)=O IFTVANMRTIHKML-WDSKDSINSA-N 0.000 description 1
- FUSPCLTUKXQREV-ACZMJKKPSA-N Ala-Glu-Ala Chemical compound [H]N[C@@H](C)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](C)C(O)=O FUSPCLTUKXQREV-ACZMJKKPSA-N 0.000 description 1
- NWVVKQZOVSTDBQ-CIUDSAMLSA-N Ala-Glu-Arg Chemical compound [H]N[C@@H](C)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CCCNC(N)=N)C(O)=O NWVVKQZOVSTDBQ-CIUDSAMLSA-N 0.000 description 1
- NJPMYXWVWQWCSR-ACZMJKKPSA-N Ala-Glu-Asn Chemical compound C[C@H](N)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CC(N)=O)C(O)=O NJPMYXWVWQWCSR-ACZMJKKPSA-N 0.000 description 1
- WKOBSJOZRJJVRZ-FXQIFTODSA-N Ala-Glu-Glu Chemical compound [H]N[C@@H](C)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CCC(O)=O)C(O)=O WKOBSJOZRJJVRZ-FXQIFTODSA-N 0.000 description 1
- UHMQKOBNPRAZGB-CIUDSAMLSA-N Ala-Glu-Met Chemical compound C[C@@H](C(=O)N[C@@H](CCC(=O)O)C(=O)N[C@@H](CCSC)C(=O)O)N UHMQKOBNPRAZGB-CIUDSAMLSA-N 0.000 description 1
- FBHOPGDGELNWRH-DRZSPHRISA-N Ala-Glu-Phe Chemical compound [H]N[C@@H](C)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CC1=CC=CC=C1)C(O)=O FBHOPGDGELNWRH-DRZSPHRISA-N 0.000 description 1
- ZVFVBBGVOILKPO-WHFBIAKZSA-N Ala-Gly-Ala Chemical compound C[C@H](N)C(=O)NCC(=O)N[C@@H](C)C(O)=O ZVFVBBGVOILKPO-WHFBIAKZSA-N 0.000 description 1
- BTBUEVAGZCKULD-XPUUQOCRSA-N Ala-Gly-His Chemical compound C[C@H](N)C(=O)NCC(=O)N[C@H](C(O)=O)CC1=CN=CN1 BTBUEVAGZCKULD-XPUUQOCRSA-N 0.000 description 1
- HUUOZYZWNCXTFK-INTQDDNPSA-N Ala-His-Pro Chemical compound C[C@@H](C(=O)N[C@@H](CC1=CN=CN1)C(=O)N2CCC[C@@H]2C(=O)O)N HUUOZYZWNCXTFK-INTQDDNPSA-N 0.000 description 1
- OKIKVSXTXVVFDV-MMWGEVLESA-N Ala-Ile-Pro Chemical compound CC[C@H](C)[C@@H](C(=O)N1CCC[C@@H]1C(=O)O)NC(=O)[C@H](C)N OKIKVSXTXVVFDV-MMWGEVLESA-N 0.000 description 1
- OYJCVIGKMXUVKB-GARJFASQSA-N Ala-Leu-Pro Chemical compound C[C@@H](C(=O)N[C@@H](CC(C)C)C(=O)N1CCC[C@@H]1C(=O)O)N OYJCVIGKMXUVKB-GARJFASQSA-N 0.000 description 1
- SOBIAADAMRHGKH-CIUDSAMLSA-N Ala-Leu-Ser Chemical compound [H]N[C@@H](C)C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](CO)C(O)=O SOBIAADAMRHGKH-CIUDSAMLSA-N 0.000 description 1
- DCVYRWFAMZFSDA-ZLUOBGJFSA-N Ala-Ser-Ala Chemical compound [H]N[C@@H](C)C(=O)N[C@@H](CO)C(=O)N[C@@H](C)C(O)=O DCVYRWFAMZFSDA-ZLUOBGJFSA-N 0.000 description 1
- YYAVDNKUWLAFCV-ACZMJKKPSA-N Ala-Ser-Gln Chemical compound [H]N[C@@H](C)C(=O)N[C@@H](CO)C(=O)N[C@@H](CCC(N)=O)C(O)=O YYAVDNKUWLAFCV-ACZMJKKPSA-N 0.000 description 1
- NHWYNIZWLJYZAG-XVYDVKMFSA-N Ala-Ser-His Chemical compound C[C@@H](C(=O)N[C@@H](CO)C(=O)N[C@@H](CC1=CN=CN1)C(=O)O)N NHWYNIZWLJYZAG-XVYDVKMFSA-N 0.000 description 1
- ARHJJAAWNWOACN-FXQIFTODSA-N Ala-Ser-Val Chemical compound [H]N[C@@H](C)C(=O)N[C@@H](CO)C(=O)N[C@@H](C(C)C)C(O)=O ARHJJAAWNWOACN-FXQIFTODSA-N 0.000 description 1
- VNFSAYFQLXPHPY-CIQUZCHMSA-N Ala-Thr-Ile Chemical compound [H]N[C@@H](C)C(=O)N[C@@H]([C@@H](C)O)C(=O)N[C@@H]([C@@H](C)CC)C(O)=O VNFSAYFQLXPHPY-CIQUZCHMSA-N 0.000 description 1
- JJHBEVZAZXZREW-LFSVMHDDSA-N Ala-Thr-Phe Chemical compound C[C@@H](O)[C@H](NC(=O)[C@H](C)N)C(=O)N[C@@H](Cc1ccccc1)C(O)=O JJHBEVZAZXZREW-LFSVMHDDSA-N 0.000 description 1
- KTXKIYXZQFWJKB-VZFHVOOUSA-N Ala-Thr-Ser Chemical compound [H]N[C@@H](C)C(=O)N[C@@H]([C@@H](C)O)C(=O)N[C@@H](CO)C(O)=O KTXKIYXZQFWJKB-VZFHVOOUSA-N 0.000 description 1
- IETUUAHKCHOQHP-KZVJFYERSA-N Ala-Thr-Val Chemical compound CC(C)[C@H](NC(=O)[C@@H](NC(=O)[C@H](C)N)[C@@H](C)O)C(O)=O IETUUAHKCHOQHP-KZVJFYERSA-N 0.000 description 1
- DEAGTWNKODHUIY-MRFFXTKBSA-N Ala-Tyr-Trp Chemical compound [H]N[C@@H](C)C(=O)N[C@@H](CC1=CC=C(O)C=C1)C(=O)N[C@@H](CC1=CNC2=C1C=CC=C2)C(O)=O DEAGTWNKODHUIY-MRFFXTKBSA-N 0.000 description 1
- YEBZNKPPOHFZJM-BPNCWPANSA-N Ala-Tyr-Val Chemical compound [H]N[C@@H](C)C(=O)N[C@@H](CC1=CC=C(O)C=C1)C(=O)N[C@@H](C(C)C)C(O)=O YEBZNKPPOHFZJM-BPNCWPANSA-N 0.000 description 1
- IYKVSFNGSWTTNZ-GUBZILKMSA-N Ala-Val-Arg Chemical compound C[C@H](N)C(=O)N[C@@H](C(C)C)C(=O)N[C@H](C(O)=O)CCCN=C(N)N IYKVSFNGSWTTNZ-GUBZILKMSA-N 0.000 description 1
- YJHKTAMKPGFJCT-NRPADANISA-N Ala-Val-Glu Chemical compound [H]N[C@@H](C)C(=O)N[C@@H](C(C)C)C(=O)N[C@@H](CCC(O)=O)C(O)=O YJHKTAMKPGFJCT-NRPADANISA-N 0.000 description 1
- REWSWYIDQIELBE-FXQIFTODSA-N Ala-Val-Ser Chemical compound [H]N[C@@H](C)C(=O)N[C@@H](C(C)C)C(=O)N[C@@H](CO)C(O)=O REWSWYIDQIELBE-FXQIFTODSA-N 0.000 description 1
- OMSKGWFGWCQFBD-KZVJFYERSA-N Ala-Val-Thr Chemical compound [H]N[C@@H](C)C(=O)N[C@@H](C(C)C)C(=O)N[C@@H]([C@@H](C)O)C(O)=O OMSKGWFGWCQFBD-KZVJFYERSA-N 0.000 description 1
- 102000002260 Alkaline Phosphatase Human genes 0.000 description 1
- 108020004774 Alkaline Phosphatase Proteins 0.000 description 1
- 241000429811 Alternariaster helianthi Species 0.000 description 1
- USFZMSVCRYTOJT-UHFFFAOYSA-N Ammonium acetate Chemical compound N.CC(O)=O USFZMSVCRYTOJT-UHFFFAOYSA-N 0.000 description 1
- 239000005695 Ammonium acetate Substances 0.000 description 1
- NLXLAEXVIDQMFP-UHFFFAOYSA-N Ammonium chloride Substances [NH4+].[Cl-] NLXLAEXVIDQMFP-UHFFFAOYSA-N 0.000 description 1
- VHUUQVKOLVNVRT-UHFFFAOYSA-N Ammonium hydroxide Chemical compound [NH4+].[OH-] VHUUQVKOLVNVRT-UHFFFAOYSA-N 0.000 description 1
- 101710170230 Antimicrobial peptide 1 Proteins 0.000 description 1
- VBFJESQBIWCWRL-DCAQKATOSA-N Arg-Ala-Lys Chemical compound NCCCC[C@@H](C(O)=O)NC(=O)[C@H](C)NC(=O)[C@@H](N)CCCNC(N)=N VBFJESQBIWCWRL-DCAQKATOSA-N 0.000 description 1
- SBVJJNJLFWSJOV-UBHSHLNASA-N Arg-Ala-Phe Chemical compound NC(=N)NCCC[C@H](N)C(=O)N[C@@H](C)C(=O)N[C@H](C(O)=O)CC1=CC=CC=C1 SBVJJNJLFWSJOV-UBHSHLNASA-N 0.000 description 1
- OTOXOKCIIQLMFH-KZVJFYERSA-N Arg-Ala-Thr Chemical compound C[C@@H](O)[C@@H](C(O)=O)NC(=O)[C@H](C)NC(=O)[C@@H](N)CCCN=C(N)N OTOXOKCIIQLMFH-KZVJFYERSA-N 0.000 description 1
- MUXONAMCEUBVGA-DCAQKATOSA-N Arg-Arg-Gln Chemical compound NC(N)=NCCC[C@H](N)C(=O)N[C@@H](CCCN=C(N)N)C(=O)N[C@@H](CCC(N)=O)C(O)=O MUXONAMCEUBVGA-DCAQKATOSA-N 0.000 description 1
- RVDVDRUZWZIBJQ-CIUDSAMLSA-N Arg-Asn-Glu Chemical compound [H]N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CC(N)=O)C(=O)N[C@@H](CCC(O)=O)C(O)=O RVDVDRUZWZIBJQ-CIUDSAMLSA-N 0.000 description 1
- NTAZNGWBXRVEDJ-FXQIFTODSA-N Arg-Asp-Asp Chemical compound [H]N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CC(O)=O)C(=O)N[C@@H](CC(O)=O)C(O)=O NTAZNGWBXRVEDJ-FXQIFTODSA-N 0.000 description 1
- ALOVURZCXKYKJC-NAKRPEOUSA-N Arg-Asp-Gln-Ser Chemical compound N[C@@H](CCCN=C(N)N)C(=O)N[C@@H](CC(O)=O)C(=O)N[C@@H](CCC(N)=O)C(=O)N[C@@H](CO)C(O)=O ALOVURZCXKYKJC-NAKRPEOUSA-N 0.000 description 1
- OZNSCVPYWZRQPY-CIUDSAMLSA-N Arg-Asp-Glu Chemical compound [H]N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CC(O)=O)C(=O)N[C@@H](CCC(O)=O)C(O)=O OZNSCVPYWZRQPY-CIUDSAMLSA-N 0.000 description 1
- HKRXJBBCQBAGIM-FXQIFTODSA-N Arg-Asp-Ser Chemical compound C(C[C@@H](C(=O)N[C@@H](CC(=O)O)C(=O)N[C@@H](CO)C(=O)O)N)CN=C(N)N HKRXJBBCQBAGIM-FXQIFTODSA-N 0.000 description 1
- FBLMOFHNVQBKRR-IHRRRGAJSA-N Arg-Asp-Tyr Chemical compound NC(N)=NCCC[C@H](N)C(=O)N[C@@H](CC(O)=O)C(=O)N[C@H](C(O)=O)CC1=CC=C(O)C=C1 FBLMOFHNVQBKRR-IHRRRGAJSA-N 0.000 description 1
- IGULQRCJLQQPSM-DCAQKATOSA-N Arg-Cys-Leu Chemical compound [H]N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CS)C(=O)N[C@@H](CC(C)C)C(O)=O IGULQRCJLQQPSM-DCAQKATOSA-N 0.000 description 1
- VNFWDYWTSHFRRG-SRVKXCTJSA-N Arg-Gln-Leu Chemical compound [H]N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CCC(N)=O)C(=O)N[C@@H](CC(C)C)C(O)=O VNFWDYWTSHFRRG-SRVKXCTJSA-N 0.000 description 1
- QAODJPUKWNNNRP-DCAQKATOSA-N Arg-Glu-Arg Chemical compound NC(N)=NCCC[C@H](N)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CCCN=C(N)N)C(O)=O QAODJPUKWNNNRP-DCAQKATOSA-N 0.000 description 1
- MZRBYBIQTIKERR-GUBZILKMSA-N Arg-Glu-Gln Chemical compound [H]N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CCC(N)=O)C(O)=O MZRBYBIQTIKERR-GUBZILKMSA-N 0.000 description 1
- OHYQKYUTLIPFOX-ZPFDUUQYSA-N Arg-Glu-Ile Chemical compound [H]N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H]([C@@H](C)CC)C(O)=O OHYQKYUTLIPFOX-ZPFDUUQYSA-N 0.000 description 1
- NXDXECQFKHXHAM-HJGDQZAQSA-N Arg-Glu-Thr Chemical compound [H]N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H]([C@@H](C)O)C(O)=O NXDXECQFKHXHAM-HJGDQZAQSA-N 0.000 description 1
- IYMAXBFPHPZYIK-BQBZGAKWSA-N Arg-Gly-Asp Chemical compound NC(N)=NCCC[C@H](N)C(=O)NCC(=O)N[C@@H](CC(O)=O)C(O)=O IYMAXBFPHPZYIK-BQBZGAKWSA-N 0.000 description 1
- WVNFNPGXYADPPO-BQBZGAKWSA-N Arg-Gly-Ser Chemical compound NC(N)=NCCC[C@H](N)C(=O)NCC(=O)N[C@@H](CO)C(O)=O WVNFNPGXYADPPO-BQBZGAKWSA-N 0.000 description 1
- VRZDJJWOFXMFRO-ZFWWWQNUSA-N Arg-Gly-Trp Chemical compound [H]N[C@@H](CCCNC(N)=N)C(=O)NCC(=O)N[C@@H](CC1=CNC2=C1C=CC=C2)C(O)=O VRZDJJWOFXMFRO-ZFWWWQNUSA-N 0.000 description 1
- JTZUZBADHGISJD-SRVKXCTJSA-N Arg-His-Glu Chemical compound [H]N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CC1=CNC=N1)C(=O)N[C@@H](CCC(O)=O)C(O)=O JTZUZBADHGISJD-SRVKXCTJSA-N 0.000 description 1
- RKQRHMKFNBYOTN-IHRRRGAJSA-N Arg-His-Lys Chemical compound C1=C(NC=N1)C[C@@H](C(=O)N[C@@H](CCCCN)C(=O)O)NC(=O)[C@H](CCCN=C(N)N)N RKQRHMKFNBYOTN-IHRRRGAJSA-N 0.000 description 1
- UPKMBGAAEZGHOC-RWMBFGLXSA-N Arg-His-Pro Chemical compound C1C[C@@H](N(C1)C(=O)[C@H](CC2=CN=CN2)NC(=O)[C@H](CCCN=C(N)N)N)C(=O)O UPKMBGAAEZGHOC-RWMBFGLXSA-N 0.000 description 1
- NVUIWHJLPSZZQC-CYDGBPFRSA-N Arg-Ile-Arg Chemical compound NC(N)=NCCC[C@H](N)C(=O)N[C@@H]([C@@H](C)CC)C(=O)N[C@@H](CCCN=C(N)N)C(O)=O NVUIWHJLPSZZQC-CYDGBPFRSA-N 0.000 description 1
- OOIMKQRCPJBGPD-XUXIUFHCSA-N Arg-Ile-Leu Chemical compound [H]N[C@@H](CCCNC(N)=N)C(=O)N[C@@H]([C@@H](C)CC)C(=O)N[C@@H](CC(C)C)C(O)=O OOIMKQRCPJBGPD-XUXIUFHCSA-N 0.000 description 1
- GNYUVVJYGJFKHN-RVMXOQNASA-N Arg-Ile-Pro Chemical compound CC[C@H](C)[C@@H](C(=O)N1CCC[C@@H]1C(=O)O)NC(=O)[C@H](CCCN=C(N)N)N GNYUVVJYGJFKHN-RVMXOQNASA-N 0.000 description 1
- OTZMRMHZCMZOJZ-SRVKXCTJSA-N Arg-Leu-Glu Chemical compound [H]N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](CCC(O)=O)C(O)=O OTZMRMHZCMZOJZ-SRVKXCTJSA-N 0.000 description 1
- UZGFHWIJWPUPOH-IHRRRGAJSA-N Arg-Leu-Lys Chemical compound CC(C)C[C@@H](C(=O)N[C@@H](CCCCN)C(=O)O)NC(=O)[C@H](CCCN=C(N)N)N UZGFHWIJWPUPOH-IHRRRGAJSA-N 0.000 description 1
- YVTHEZNOKSAWRW-DCAQKATOSA-N Arg-Lys-Ala Chemical compound [H]N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CCCCN)C(=O)N[C@@H](C)C(O)=O YVTHEZNOKSAWRW-DCAQKATOSA-N 0.000 description 1
- NGTYEHIRESTSRX-UWVGGRQHSA-N Arg-Lys-Gly Chemical compound NCCCC[C@@H](C(=O)NCC(O)=O)NC(=O)[C@@H](N)CCCN=C(N)N NGTYEHIRESTSRX-UWVGGRQHSA-N 0.000 description 1
- CLICCYPMVFGUOF-IHRRRGAJSA-N Arg-Lys-Leu Chemical compound [H]N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CCCCN)C(=O)N[C@@H](CC(C)C)C(O)=O CLICCYPMVFGUOF-IHRRRGAJSA-N 0.000 description 1
- NPAVRDPEFVKELR-DCAQKATOSA-N Arg-Lys-Ser Chemical compound [H]N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CCCCN)C(=O)N[C@@H](CO)C(O)=O NPAVRDPEFVKELR-DCAQKATOSA-N 0.000 description 1
- JOADBFCFJGNIKF-GUBZILKMSA-N Arg-Met-Ala Chemical compound [H]N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CCSC)C(=O)N[C@@H](C)C(O)=O JOADBFCFJGNIKF-GUBZILKMSA-N 0.000 description 1
- NYDIVDKTULRINZ-AVGNSLFASA-N Arg-Met-Lys Chemical compound CSCC[C@@H](C(=O)N[C@@H](CCCCN)C(=O)O)NC(=O)[C@H](CCCN=C(N)N)N NYDIVDKTULRINZ-AVGNSLFASA-N 0.000 description 1
- KSUALAGYYLQSHJ-RCWTZXSCSA-N Arg-Met-Thr Chemical compound [H]N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CCSC)C(=O)N[C@@H]([C@@H](C)O)C(O)=O KSUALAGYYLQSHJ-RCWTZXSCSA-N 0.000 description 1
- CZUHPNLXLWMYMG-UBHSHLNASA-N Arg-Phe-Ala Chemical compound NC(N)=NCCC[C@H](N)C(=O)N[C@H](C(=O)N[C@@H](C)C(O)=O)CC1=CC=CC=C1 CZUHPNLXLWMYMG-UBHSHLNASA-N 0.000 description 1
- INXWADWANGLMPJ-JYJNAYRXSA-N Arg-Phe-Arg Chemical compound NC(=N)NCCC[C@H](N)C(=O)N[C@H](C(=O)N[C@@H](CCCNC(N)=N)C(O)=O)CC1=CC=CC=C1 INXWADWANGLMPJ-JYJNAYRXSA-N 0.000 description 1
- YTMKMRSYXHBGER-IHRRRGAJSA-N Arg-Phe-Asn Chemical compound C1=CC=C(C=C1)C[C@@H](C(=O)N[C@@H](CC(=O)N)C(=O)O)NC(=O)[C@H](CCCN=C(N)N)N YTMKMRSYXHBGER-IHRRRGAJSA-N 0.000 description 1
- BSYKSCBTTQKOJG-GUBZILKMSA-N Arg-Pro-Ala Chemical compound [H]N[C@@H](CCCNC(N)=N)C(=O)N1CCC[C@H]1C(=O)N[C@@H](C)C(O)=O BSYKSCBTTQKOJG-GUBZILKMSA-N 0.000 description 1
- YFHATWYGAAXQCF-JYJNAYRXSA-N Arg-Pro-Phe Chemical compound NC(N)=NCCC[C@H](N)C(=O)N1CCC[C@H]1C(=O)N[C@H](C(O)=O)CC1=CC=CC=C1 YFHATWYGAAXQCF-JYJNAYRXSA-N 0.000 description 1
- YCYXHLZRUSJITQ-SRVKXCTJSA-N Arg-Pro-Pro Chemical compound NC(=N)NCCC[C@H](N)C(=O)N1CCC[C@H]1C(=O)N1[C@H](C(O)=O)CCC1 YCYXHLZRUSJITQ-SRVKXCTJSA-N 0.000 description 1
- VRTWYUYCJGNFES-CIUDSAMLSA-N Arg-Ser-Gln Chemical compound [H]N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CO)C(=O)N[C@@H](CCC(N)=O)C(O)=O VRTWYUYCJGNFES-CIUDSAMLSA-N 0.000 description 1
- BECXEHHOZNFFFX-IHRRRGAJSA-N Arg-Ser-Tyr Chemical compound [H]N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CO)C(=O)N[C@@H](CC1=CC=C(O)C=C1)C(O)=O BECXEHHOZNFFFX-IHRRRGAJSA-N 0.000 description 1
- PYDIIVKGTBRIEL-SZMVWBNQSA-N Arg-Trp-Pro Chemical compound [H]N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CC1=CNC2=C1C=CC=C2)C(=O)N1CCC[C@H]1C(O)=O PYDIIVKGTBRIEL-SZMVWBNQSA-N 0.000 description 1
- IZSMEUDYADKZTJ-KJEVXHAQSA-N Arg-Tyr-Thr Chemical compound [H]N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CC1=CC=C(O)C=C1)C(=O)N[C@@H]([C@@H](C)O)C(O)=O IZSMEUDYADKZTJ-KJEVXHAQSA-N 0.000 description 1
- CPTXATAOUQJQRO-GUBZILKMSA-N Arg-Val-Ser Chemical compound [H]N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](C(C)C)C(=O)N[C@@H](CO)C(O)=O CPTXATAOUQJQRO-GUBZILKMSA-N 0.000 description 1
- LEFKSBYHUGUWLP-ACZMJKKPSA-N Asn-Ala-Glu Chemical compound [H]N[C@@H](CC(N)=O)C(=O)N[C@@H](C)C(=O)N[C@@H](CCC(O)=O)C(O)=O LEFKSBYHUGUWLP-ACZMJKKPSA-N 0.000 description 1
- AYZAWXAPBAYCHO-CIUDSAMLSA-N Asn-Asn-His Chemical compound C1=C(NC=N1)C[C@@H](C(=O)O)NC(=O)[C@H](CC(=O)N)NC(=O)[C@H](CC(=O)N)N AYZAWXAPBAYCHO-CIUDSAMLSA-N 0.000 description 1
- APHUDFFMXFYRKP-CIUDSAMLSA-N Asn-Asn-Lys Chemical compound C(CCN)C[C@@H](C(=O)O)NC(=O)[C@H](CC(=O)N)NC(=O)[C@H](CC(=O)N)N APHUDFFMXFYRKP-CIUDSAMLSA-N 0.000 description 1
- NVGWESORMHFISY-SRVKXCTJSA-N Asn-Asn-Phe Chemical compound [H]N[C@@H](CC(N)=O)C(=O)N[C@@H](CC(N)=O)C(=O)N[C@@H](CC1=CC=CC=C1)C(O)=O NVGWESORMHFISY-SRVKXCTJSA-N 0.000 description 1
- NLCDVZJDEXIDDL-BIIVOSGPSA-N Asn-Asn-Pro Chemical compound C1C[C@@H](N(C1)C(=O)[C@H](CC(=O)N)NC(=O)[C@H](CC(=O)N)N)C(=O)O NLCDVZJDEXIDDL-BIIVOSGPSA-N 0.000 description 1
- PAXHINASXXXILC-SRVKXCTJSA-N Asn-Asp-Tyr Chemical compound C1=CC(=CC=C1C[C@@H](C(=O)O)NC(=O)[C@H](CC(=O)O)NC(=O)[C@H](CC(=O)N)N)O PAXHINASXXXILC-SRVKXCTJSA-N 0.000 description 1
- XWFPGQVLOVGSLU-CIUDSAMLSA-N Asn-Gln-Arg Chemical compound NC(=O)C[C@H](N)C(=O)N[C@@H](CCC(N)=O)C(=O)N[C@H](C(O)=O)CCCN=C(N)N XWFPGQVLOVGSLU-CIUDSAMLSA-N 0.000 description 1
- AYKKKGFJXIDYLX-ACZMJKKPSA-N Asn-Gln-Asn Chemical compound [H]N[C@@H](CC(N)=O)C(=O)N[C@@H](CCC(N)=O)C(=O)N[C@@H](CC(N)=O)C(O)=O AYKKKGFJXIDYLX-ACZMJKKPSA-N 0.000 description 1
- HJRBIWRXULGMOA-ACZMJKKPSA-N Asn-Gln-Asp Chemical compound [H]N[C@@H](CC(N)=O)C(=O)N[C@@H](CCC(N)=O)C(=O)N[C@@H](CC(O)=O)C(O)=O HJRBIWRXULGMOA-ACZMJKKPSA-N 0.000 description 1
- KWQPAXYXVMHJJR-AVGNSLFASA-N Asn-Gln-Tyr Chemical compound NC(=O)C[C@H](N)C(=O)N[C@@H](CCC(N)=O)C(=O)N[C@H](C(O)=O)CC1=CC=C(O)C=C1 KWQPAXYXVMHJJR-AVGNSLFASA-N 0.000 description 1
- PPMTUXJSQDNUDE-CIUDSAMLSA-N Asn-Glu-Arg Chemical compound NC(=O)C[C@H](N)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@H](C(O)=O)CCCN=C(N)N PPMTUXJSQDNUDE-CIUDSAMLSA-N 0.000 description 1
- MSBDSTRUMZFSEU-PEFMBERDSA-N Asn-Glu-Ile Chemical compound [H]N[C@@H](CC(N)=O)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H]([C@@H](C)CC)C(O)=O MSBDSTRUMZFSEU-PEFMBERDSA-N 0.000 description 1
- DDPXDCKYWDGZAL-BQBZGAKWSA-N Asn-Gly-Arg Chemical compound NC(=O)C[C@H](N)C(=O)NCC(=O)N[C@H](C(O)=O)CCCN=C(N)N DDPXDCKYWDGZAL-BQBZGAKWSA-N 0.000 description 1
- OLVIPTLKNSAYRJ-YUMQZZPRSA-N Asn-Gly-Lys Chemical compound C(CCN)C[C@@H](C(=O)O)NC(=O)CNC(=O)[C@H](CC(=O)N)N OLVIPTLKNSAYRJ-YUMQZZPRSA-N 0.000 description 1
- UDSVWSUXKYXSTR-QWRGUYRKSA-N Asn-Gly-Tyr Chemical compound [H]N[C@@H](CC(N)=O)C(=O)NCC(=O)N[C@@H](CC1=CC=C(O)C=C1)C(O)=O UDSVWSUXKYXSTR-QWRGUYRKSA-N 0.000 description 1
- ZKDGORKGHPCZOV-DCAQKATOSA-N Asn-His-Arg Chemical compound C1=C(NC=N1)C[C@@H](C(=O)N[C@@H](CCCN=C(N)N)C(=O)O)NC(=O)[C@H](CC(=O)N)N ZKDGORKGHPCZOV-DCAQKATOSA-N 0.000 description 1
- QEQVUHQQYDZUEN-GUBZILKMSA-N Asn-His-Glu Chemical compound C1=C(NC=N1)C[C@@H](C(=O)N[C@@H](CCC(=O)O)C(=O)O)NC(=O)[C@H](CC(=O)N)N QEQVUHQQYDZUEN-GUBZILKMSA-N 0.000 description 1
- XVBDDUPJVQXDSI-PEFMBERDSA-N Asn-Ile-Glu Chemical compound CC[C@H](C)[C@@H](C(=O)N[C@@H](CCC(=O)O)C(=O)O)NC(=O)[C@H](CC(=O)N)N XVBDDUPJVQXDSI-PEFMBERDSA-N 0.000 description 1
- GLWFAWNYGWBMOC-SRVKXCTJSA-N Asn-Leu-Leu Chemical compound [H]N[C@@H](CC(N)=O)C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](CC(C)C)C(O)=O GLWFAWNYGWBMOC-SRVKXCTJSA-N 0.000 description 1
- FODVBOKTYKYRFJ-CIUDSAMLSA-N Asn-Lys-Cys Chemical compound C(CCN)C[C@@H](C(=O)N[C@@H](CS)C(=O)O)NC(=O)[C@H](CC(=O)N)N FODVBOKTYKYRFJ-CIUDSAMLSA-N 0.000 description 1
- ORJQQZIXTOYGGH-SRVKXCTJSA-N Asn-Lys-Leu Chemical compound [H]N[C@@H](CC(N)=O)C(=O)N[C@@H](CCCCN)C(=O)N[C@@H](CC(C)C)C(O)=O ORJQQZIXTOYGGH-SRVKXCTJSA-N 0.000 description 1
- OROMFUQQTSWUTI-IHRRRGAJSA-N Asn-Phe-Arg Chemical compound C1=CC=C(C=C1)C[C@@H](C(=O)N[C@@H](CCCN=C(N)N)C(=O)O)NC(=O)[C@H](CC(=O)N)N OROMFUQQTSWUTI-IHRRRGAJSA-N 0.000 description 1
- HZZIFFOVHLWGCS-KKUMJFAQSA-N Asn-Phe-Leu Chemical compound [H]N[C@@H](CC(N)=O)C(=O)N[C@@H](CC1=CC=CC=C1)C(=O)N[C@@H](CC(C)C)C(O)=O HZZIFFOVHLWGCS-KKUMJFAQSA-N 0.000 description 1
- ZJIFRAPZHAGLGR-MELADBBJSA-N Asn-Phe-Pro Chemical compound C1C[C@@H](N(C1)C(=O)[C@H](CC2=CC=CC=C2)NC(=O)[C@H](CC(=O)N)N)C(=O)O ZJIFRAPZHAGLGR-MELADBBJSA-N 0.000 description 1
- RBOBTTLFPRSXKZ-BZSNNMDCSA-N Asn-Phe-Tyr Chemical compound [H]N[C@@H](CC(N)=O)C(=O)N[C@@H](CC1=CC=CC=C1)C(=O)N[C@@H](CC1=CC=C(O)C=C1)C(O)=O RBOBTTLFPRSXKZ-BZSNNMDCSA-N 0.000 description 1
- QXOPPIDJKPEKCW-GUBZILKMSA-N Asn-Pro-Arg Chemical compound C1C[C@H](N(C1)C(=O)[C@H](CC(=O)N)N)C(=O)N[C@@H](CCCN=C(N)N)C(=O)O QXOPPIDJKPEKCW-GUBZILKMSA-N 0.000 description 1
- JTXVXGXTRXMOFJ-FXQIFTODSA-N Asn-Pro-Asn Chemical compound NC(=O)C[C@H](N)C(=O)N1CCC[C@H]1C(=O)N[C@@H](CC(N)=O)C(O)=O JTXVXGXTRXMOFJ-FXQIFTODSA-N 0.000 description 1
- YRTOMUMWSTUQAX-FXQIFTODSA-N Asn-Pro-Asp Chemical compound NC(=O)C[C@H](N)C(=O)N1CCC[C@H]1C(=O)N[C@@H](CC(O)=O)C(O)=O YRTOMUMWSTUQAX-FXQIFTODSA-N 0.000 description 1
- GKKUBLFXKRDMFC-BQBZGAKWSA-N Asn-Pro-Gly Chemical compound [H]N[C@@H](CC(N)=O)C(=O)N1CCC[C@H]1C(=O)NCC(O)=O GKKUBLFXKRDMFC-BQBZGAKWSA-N 0.000 description 1
- NJSNXIOKBHPFMB-GMOBBJLQSA-N Asn-Pro-Ile Chemical compound CC[C@H](C)[C@@H](C(=O)O)NC(=O)[C@@H]1CCCN1C(=O)[C@H](CC(=O)N)N NJSNXIOKBHPFMB-GMOBBJLQSA-N 0.000 description 1
- UGXYFDQFLVCDFC-CIUDSAMLSA-N Asn-Ser-Lys Chemical compound NCCCC[C@@H](C(O)=O)NC(=O)[C@H](CO)NC(=O)[C@@H](N)CC(N)=O UGXYFDQFLVCDFC-CIUDSAMLSA-N 0.000 description 1
- HPASIOLTWSNMFB-OLHMAJIHSA-N Asn-Thr-Asp Chemical compound [H]N[C@@H](CC(N)=O)C(=O)N[C@@H]([C@@H](C)O)C(=O)N[C@@H](CC(O)=O)C(O)=O HPASIOLTWSNMFB-OLHMAJIHSA-N 0.000 description 1
- QUMKPKWYDVMGNT-NUMRIWBASA-N Asn-Thr-Gln Chemical compound C[C@H]([C@@H](C(=O)N[C@@H](CCC(=O)N)C(=O)O)NC(=O)[C@H](CC(=O)N)N)O QUMKPKWYDVMGNT-NUMRIWBASA-N 0.000 description 1
- QIRJQYQOIKBPBZ-IHRRRGAJSA-N Asn-Tyr-Arg Chemical compound [H]N[C@@H](CC(N)=O)C(=O)N[C@@H](CC1=CC=C(O)C=C1)C(=O)N[C@@H](CCCNC(N)=N)C(O)=O QIRJQYQOIKBPBZ-IHRRRGAJSA-N 0.000 description 1
- VTYQAQFKMQTKQD-ACZMJKKPSA-N Asp-Ala-Gln Chemical compound [H]N[C@@H](CC(O)=O)C(=O)N[C@@H](C)C(=O)N[C@@H](CCC(N)=O)C(O)=O VTYQAQFKMQTKQD-ACZMJKKPSA-N 0.000 description 1
- SLHOOKXYTYAJGQ-XVYDVKMFSA-N Asp-Ala-His Chemical compound OC(=O)C[C@H](N)C(=O)N[C@@H](C)C(=O)N[C@H](C(O)=O)CC1=CNC=N1 SLHOOKXYTYAJGQ-XVYDVKMFSA-N 0.000 description 1
- XBQSLMACWDXWLJ-GHCJXIJMSA-N Asp-Ala-Ile Chemical compound [H]N[C@@H](CC(O)=O)C(=O)N[C@@H](C)C(=O)N[C@@H]([C@@H](C)CC)C(O)=O XBQSLMACWDXWLJ-GHCJXIJMSA-N 0.000 description 1
- ZLGKHJHFYSRUBH-FXQIFTODSA-N Asp-Arg-Asp Chemical compound [H]N[C@@H](CC(O)=O)C(=O)N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CC(O)=O)C(O)=O ZLGKHJHFYSRUBH-FXQIFTODSA-N 0.000 description 1
- AXXCUABIFZPKPM-BQBZGAKWSA-N Asp-Arg-Gly Chemical compound [H]N[C@@H](CC(O)=O)C(=O)N[C@@H](CCCNC(N)=N)C(=O)NCC(O)=O AXXCUABIFZPKPM-BQBZGAKWSA-N 0.000 description 1
- FAEIQWHBRBWUBN-FXQIFTODSA-N Asp-Arg-Ser Chemical compound C(C[C@@H](C(=O)N[C@@H](CO)C(=O)O)NC(=O)[C@H](CC(=O)O)N)CN=C(N)N FAEIQWHBRBWUBN-FXQIFTODSA-N 0.000 description 1
- UQBGYPFHWFZMCD-ZLUOBGJFSA-N Asp-Asn-Asn Chemical compound OC(=O)C[C@H](N)C(=O)N[C@@H](CC(N)=O)C(=O)N[C@@H](CC(N)=O)C(O)=O UQBGYPFHWFZMCD-ZLUOBGJFSA-N 0.000 description 1
- ATYWBXGNXZYZGI-ACZMJKKPSA-N Asp-Asn-Gln Chemical compound [H]N[C@@H](CC(O)=O)C(=O)N[C@@H](CC(N)=O)C(=O)N[C@@H](CCC(N)=O)C(O)=O ATYWBXGNXZYZGI-ACZMJKKPSA-N 0.000 description 1
- ZELQAFZSJOBEQS-ACZMJKKPSA-N Asp-Asn-Glu Chemical compound [H]N[C@@H](CC(O)=O)C(=O)N[C@@H](CC(N)=O)C(=O)N[C@@H](CCC(O)=O)C(O)=O ZELQAFZSJOBEQS-ACZMJKKPSA-N 0.000 description 1
- JDHOJQJMWBKHDB-CIUDSAMLSA-N Asp-Asn-Lys Chemical compound C(CCN)C[C@@H](C(=O)O)NC(=O)[C@H](CC(=O)N)NC(=O)[C@H](CC(=O)O)N JDHOJQJMWBKHDB-CIUDSAMLSA-N 0.000 description 1
- RDRMWJBLOSRRAW-BYULHYEWSA-N Asp-Asn-Val Chemical compound [H]N[C@@H](CC(O)=O)C(=O)N[C@@H](CC(N)=O)C(=O)N[C@@H](C(C)C)C(O)=O RDRMWJBLOSRRAW-BYULHYEWSA-N 0.000 description 1
- JGDBHIVECJGXJA-FXQIFTODSA-N Asp-Asp-Arg Chemical compound [H]N[C@@H](CC(O)=O)C(=O)N[C@@H](CC(O)=O)C(=O)N[C@@H](CCCNC(N)=N)C(O)=O JGDBHIVECJGXJA-FXQIFTODSA-N 0.000 description 1
- QOVWVLLHMMCFFY-ZLUOBGJFSA-N Asp-Asp-Asn Chemical compound [H]N[C@@H](CC(O)=O)C(=O)N[C@@H](CC(O)=O)C(=O)N[C@@H](CC(N)=O)C(O)=O QOVWVLLHMMCFFY-ZLUOBGJFSA-N 0.000 description 1
- WCFCYFDBMNFSPA-ACZMJKKPSA-N Asp-Asp-Glu Chemical compound OC(=O)C[C@H](N)C(=O)N[C@@H](CC(O)=O)C(=O)N[C@H](C(O)=O)CCC(O)=O WCFCYFDBMNFSPA-ACZMJKKPSA-N 0.000 description 1
- SBHUBSDEZQFJHJ-CIUDSAMLSA-N Asp-Asp-Leu Chemical compound CC(C)C[C@@H](C(O)=O)NC(=O)[C@H](CC(O)=O)NC(=O)[C@@H](N)CC(O)=O SBHUBSDEZQFJHJ-CIUDSAMLSA-N 0.000 description 1
- QXHVOUSPVAWEMX-ZLUOBGJFSA-N Asp-Asp-Ser Chemical compound OC(=O)C[C@H](N)C(=O)N[C@@H](CC(O)=O)C(=O)N[C@@H](CO)C(O)=O QXHVOUSPVAWEMX-ZLUOBGJFSA-N 0.000 description 1
- APYNREQHZOGYHV-ACZMJKKPSA-N Asp-Cys-Gln Chemical compound C(CC(=O)N)[C@@H](C(=O)O)NC(=O)[C@H](CS)NC(=O)[C@H](CC(=O)O)N APYNREQHZOGYHV-ACZMJKKPSA-N 0.000 description 1
- RSMIHCFQDCVVBR-CIUDSAMLSA-N Asp-Gln-Arg Chemical compound OC(=O)C[C@H](N)C(=O)N[C@@H](CCC(N)=O)C(=O)N[C@H](C(O)=O)CCCNC(N)=N RSMIHCFQDCVVBR-CIUDSAMLSA-N 0.000 description 1
- IJHUZMGJRGNXIW-CIUDSAMLSA-N Asp-Glu-Arg Chemical compound [H]N[C@@H](CC(O)=O)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CCCNC(N)=N)C(O)=O IJHUZMGJRGNXIW-CIUDSAMLSA-N 0.000 description 1
- VFUXXFVCYZPOQG-WDSKDSINSA-N Asp-Glu-Gly Chemical compound [H]N[C@@H](CC(O)=O)C(=O)N[C@@H](CCC(O)=O)C(=O)NCC(O)=O VFUXXFVCYZPOQG-WDSKDSINSA-N 0.000 description 1
- PDECQIHABNQRHN-GUBZILKMSA-N Asp-Glu-Leu Chemical compound CC(C)C[C@@H](C(O)=O)NC(=O)[C@H](CCC(O)=O)NC(=O)[C@@H](N)CC(O)=O PDECQIHABNQRHN-GUBZILKMSA-N 0.000 description 1
- VIRHEUMYXXLCBF-WDSKDSINSA-N Asp-Gly-Glu Chemical compound [H]N[C@@H](CC(O)=O)C(=O)NCC(=O)N[C@@H](CCC(O)=O)C(O)=O VIRHEUMYXXLCBF-WDSKDSINSA-N 0.000 description 1
- WSGVTKZFVJSJOG-RCOVLWMOSA-N Asp-Gly-Val Chemical compound [H]N[C@@H](CC(O)=O)C(=O)NCC(=O)N[C@@H](C(C)C)C(O)=O WSGVTKZFVJSJOG-RCOVLWMOSA-N 0.000 description 1
- QHHVSXGWLYEAGX-GUBZILKMSA-N Asp-His-Gln Chemical compound C1=C(NC=N1)C[C@@H](C(=O)N[C@@H](CCC(=O)N)C(=O)O)NC(=O)[C@H](CC(=O)O)N QHHVSXGWLYEAGX-GUBZILKMSA-N 0.000 description 1
- SPWXXPFDTMYTRI-IUKAMOBKSA-N Asp-Ile-Thr Chemical compound [H]N[C@@H](CC(O)=O)C(=O)N[C@@H]([C@@H](C)CC)C(=O)N[C@@H]([C@@H](C)O)C(O)=O SPWXXPFDTMYTRI-IUKAMOBKSA-N 0.000 description 1
- PAYPSKIBMDHZPI-CIUDSAMLSA-N Asp-Leu-Asp Chemical compound [H]N[C@@H](CC(O)=O)C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](CC(O)=O)C(O)=O PAYPSKIBMDHZPI-CIUDSAMLSA-N 0.000 description 1
- RQHLMGCXCZUOGT-ZPFDUUQYSA-N Asp-Leu-Ile Chemical compound [H]N[C@@H](CC(O)=O)C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H]([C@@H](C)CC)C(O)=O RQHLMGCXCZUOGT-ZPFDUUQYSA-N 0.000 description 1
- CTWCFPWFIGRAEP-CIUDSAMLSA-N Asp-Lys-Asp Chemical compound [H]N[C@@H](CC(O)=O)C(=O)N[C@@H](CCCCN)C(=O)N[C@@H](CC(O)=O)C(O)=O CTWCFPWFIGRAEP-CIUDSAMLSA-N 0.000 description 1
- WDMNFNXKGSLIOB-GUBZILKMSA-N Asp-Met-Met Chemical compound CSCC[C@@H](C(=O)N[C@@H](CCSC)C(=O)O)NC(=O)[C@H](CC(=O)O)N WDMNFNXKGSLIOB-GUBZILKMSA-N 0.000 description 1
- GPPIDDWYKJPRES-YDHLFZDLSA-N Asp-Phe-Val Chemical compound [H]N[C@@H](CC(O)=O)C(=O)N[C@@H](CC1=CC=CC=C1)C(=O)N[C@@H](C(C)C)C(O)=O GPPIDDWYKJPRES-YDHLFZDLSA-N 0.000 description 1
- ZKAOJVJQGVUIIU-GUBZILKMSA-N Asp-Pro-Arg Chemical compound OC(=O)C[C@H](N)C(=O)N1CCC[C@H]1C(=O)N[C@@H](CCCNC(N)=N)C(O)=O ZKAOJVJQGVUIIU-GUBZILKMSA-N 0.000 description 1
- BKOIIURTQAJHAT-GUBZILKMSA-N Asp-Pro-Pro Chemical compound OC(=O)C[C@H](N)C(=O)N1CCC[C@H]1C(=O)N1[C@H](C(O)=O)CCC1 BKOIIURTQAJHAT-GUBZILKMSA-N 0.000 description 1
- HCOQNGIHSXICCB-IHRRRGAJSA-N Asp-Tyr-Arg Chemical compound N[C@@H](CC(=O)O)C(=O)N[C@@H](CC1=CC=C(C=C1)O)C(=O)N[C@@H](CCCNC(N)=N)C(=O)O HCOQNGIHSXICCB-IHRRRGAJSA-N 0.000 description 1
- USENATHVGFXRNO-SRVKXCTJSA-N Asp-Tyr-Asp Chemical compound OC(=O)C[C@H](N)C(=O)N[C@H](C(=O)N[C@@H](CC(O)=O)C(O)=O)CC1=CC=C(O)C=C1 USENATHVGFXRNO-SRVKXCTJSA-N 0.000 description 1
- XWKBWZXGNXTDKY-ZKWXMUAHSA-N Asp-Val-Ala Chemical compound OC(=O)[C@H](C)NC(=O)[C@H](C(C)C)NC(=O)[C@@H](N)CC(O)=O XWKBWZXGNXTDKY-ZKWXMUAHSA-N 0.000 description 1
- GFYOIYJJMSHLSN-QXEWZRGKSA-N Asp-Val-Arg Chemical compound OC(=O)C[C@H](N)C(=O)N[C@@H](C(C)C)C(=O)N[C@@H](CCCN=C(N)N)C(O)=O GFYOIYJJMSHLSN-QXEWZRGKSA-N 0.000 description 1
- UXRVDHVARNBOIO-QSFUFRPTSA-N Asp-Val-Ile Chemical compound CC[C@H](C)[C@@H](C(=O)O)NC(=O)[C@H](C(C)C)NC(=O)[C@H](CC(=O)O)N UXRVDHVARNBOIO-QSFUFRPTSA-N 0.000 description 1
- QOJJMJKTMKNFEF-ZKWXMUAHSA-N Asp-Val-Ser Chemical compound OC[C@@H](C(O)=O)NC(=O)[C@H](C(C)C)NC(=O)[C@@H](N)CC(O)=O QOJJMJKTMKNFEF-ZKWXMUAHSA-N 0.000 description 1
- DCXYFEDJOCDNAF-UHFFFAOYSA-N Asparagine Natural products OC(=O)C(N)CC(N)=O DCXYFEDJOCDNAF-UHFFFAOYSA-N 0.000 description 1
- 241001534796 Banksia canei Species 0.000 description 1
- 241000892129 Banksia formosa Species 0.000 description 1
- 241001534840 Banksia robur Species 0.000 description 1
- UXVMQQNJUSDDNG-UHFFFAOYSA-L Calcium chloride Chemical compound [Cl-].[Cl-].[Ca+2] UXVMQQNJUSDDNG-UHFFFAOYSA-L 0.000 description 1
- 241000222122 Candida albicans Species 0.000 description 1
- 241000283707 Capra Species 0.000 description 1
- 241001157784 Cercospora nicotianae Species 0.000 description 1
- 241001136168 Clavibacter michiganensis Species 0.000 description 1
- 229910021580 Cobalt(II) chloride Inorganic materials 0.000 description 1
- 108700010070 Codon Usage Proteins 0.000 description 1
- 108020004635 Complementary DNA Proteins 0.000 description 1
- XMTDCXXLDZKAGI-ACZMJKKPSA-N Cys-Ala-Gln Chemical compound C[C@@H](C(=O)N[C@@H](CCC(=O)N)C(=O)O)NC(=O)[C@H](CS)N XMTDCXXLDZKAGI-ACZMJKKPSA-N 0.000 description 1
- UKVGHFORADMBEN-GUBZILKMSA-N Cys-Arg-Arg Chemical compound [H]N[C@@H](CS)C(=O)N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CCCNC(N)=N)C(O)=O UKVGHFORADMBEN-GUBZILKMSA-N 0.000 description 1
- CLDCTNHPILWQCW-CIUDSAMLSA-N Cys-Arg-Glu Chemical compound C(C[C@@H](C(=O)N[C@@H](CCC(=O)O)C(=O)O)NC(=O)[C@H](CS)N)CN=C(N)N CLDCTNHPILWQCW-CIUDSAMLSA-N 0.000 description 1
- UPJGYXRAPJWIHD-CIUDSAMLSA-N Cys-Asn-Leu Chemical compound [H]N[C@@H](CS)C(=O)N[C@@H](CC(N)=O)C(=O)N[C@@H](CC(C)C)C(O)=O UPJGYXRAPJWIHD-CIUDSAMLSA-N 0.000 description 1
- CPTUXCUWQIBZIF-ZLUOBGJFSA-N Cys-Asn-Ser Chemical compound SC[C@H](N)C(=O)N[C@@H](CC(N)=O)C(=O)N[C@@H](CO)C(O)=O CPTUXCUWQIBZIF-ZLUOBGJFSA-N 0.000 description 1
- VNLYIYOYUNGURO-ZLUOBGJFSA-N Cys-Asp-Ala Chemical compound C[C@@H](C(=O)O)NC(=O)[C@H](CC(=O)O)NC(=O)[C@H](CS)N VNLYIYOYUNGURO-ZLUOBGJFSA-N 0.000 description 1
- MBILEVLLOHJZMG-FXQIFTODSA-N Cys-Gln-Glu Chemical compound C(CC(=O)N)[C@@H](C(=O)N[C@@H](CCC(=O)O)C(=O)O)NC(=O)[C@H](CS)N MBILEVLLOHJZMG-FXQIFTODSA-N 0.000 description 1
- YZKOXEJTLWZOQL-GUBZILKMSA-N Cys-Gln-Leu Chemical compound CC(C)C[C@@H](C(=O)O)NC(=O)[C@H](CCC(=O)N)NC(=O)[C@H](CS)N YZKOXEJTLWZOQL-GUBZILKMSA-N 0.000 description 1
- LMXOUGMSGHFLRX-CIUDSAMLSA-N Cys-Gln-Met Chemical compound CSCC[C@@H](C(=O)O)NC(=O)[C@H](CCC(=O)N)NC(=O)[C@H](CS)N LMXOUGMSGHFLRX-CIUDSAMLSA-N 0.000 description 1
- FIADUEYFRSCCIK-CIUDSAMLSA-N Cys-Glu-Arg Chemical compound [H]N[C@@H](CS)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CCCNC(N)=N)C(O)=O FIADUEYFRSCCIK-CIUDSAMLSA-N 0.000 description 1
- ZEXHDOQQYZKOIB-ACZMJKKPSA-N Cys-Glu-Ser Chemical compound [H]N[C@@H](CS)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CO)C(O)=O ZEXHDOQQYZKOIB-ACZMJKKPSA-N 0.000 description 1
- XLLSMEFANRROJE-GUBZILKMSA-N Cys-Leu-Glu Chemical compound CC(C)C[C@@H](C(=O)N[C@@H](CCC(=O)O)C(=O)O)NC(=O)[C@H](CS)N XLLSMEFANRROJE-GUBZILKMSA-N 0.000 description 1
- UCSXXFRXHGUXCQ-SRVKXCTJSA-N Cys-Leu-Lys Chemical compound CC(C)C[C@@H](C(=O)N[C@@H](CCCCN)C(=O)O)NC(=O)[C@H](CS)N UCSXXFRXHGUXCQ-SRVKXCTJSA-N 0.000 description 1
- NIXHTNJAGGFBAW-CIUDSAMLSA-N Cys-Lys-Ser Chemical compound C(CCN)C[C@@H](C(=O)N[C@@H](CO)C(=O)O)NC(=O)[C@H](CS)N NIXHTNJAGGFBAW-CIUDSAMLSA-N 0.000 description 1
- CMYVIUWVYHOLRD-ZLUOBGJFSA-N Cys-Ser-Ala Chemical compound [H]N[C@@H](CS)C(=O)N[C@@H](CO)C(=O)N[C@@H](C)C(O)=O CMYVIUWVYHOLRD-ZLUOBGJFSA-N 0.000 description 1
- WKKKNGNJDGATNS-QEJZJMRPSA-N Cys-Trp-Glu Chemical compound [H]N[C@@H](CS)C(=O)N[C@@H](CC1=CNC2=C1C=CC=C2)C(=O)N[C@@H](CCC(O)=O)C(O)=O WKKKNGNJDGATNS-QEJZJMRPSA-N 0.000 description 1
- UGPCUUWZXRMCIJ-KKUMJFAQSA-N Cys-Tyr-Leu Chemical compound CC(C)C[C@@H](C(=O)O)NC(=O)[C@H](CC1=CC=C(C=C1)O)NC(=O)[C@H](CS)N UGPCUUWZXRMCIJ-KKUMJFAQSA-N 0.000 description 1
- ALTQTAKGRFLRLR-GUBZILKMSA-N Cys-Val-Val Chemical compound CC(C)[C@@H](C(=O)N[C@@H](C(C)C)C(=O)O)NC(=O)[C@H](CS)N ALTQTAKGRFLRLR-GUBZILKMSA-N 0.000 description 1
- 108010025905 Cystine-Knot Miniproteins Proteins 0.000 description 1
- 102000012410 DNA Ligases Human genes 0.000 description 1
- 108010061982 DNA Ligases Proteins 0.000 description 1
- 102000004594 DNA Polymerase I Human genes 0.000 description 1
- 108010017826 DNA Polymerase I Proteins 0.000 description 1
- 108010054576 Deoxyribonuclease EcoRI Proteins 0.000 description 1
- BWGNESOTFCXPMA-UHFFFAOYSA-N Dihydrogen disulfide Chemical compound SS BWGNESOTFCXPMA-UHFFFAOYSA-N 0.000 description 1
- 102100031780 Endonuclease Human genes 0.000 description 1
- 102000004190 Enzymes Human genes 0.000 description 1
- 108090000790 Enzymes Proteins 0.000 description 1
- 241000672609 Escherichia coli BL21 Species 0.000 description 1
- 241000701959 Escherichia virus Lambda Species 0.000 description 1
- 108700028146 Genetic Enhancer Elements Proteins 0.000 description 1
- DTCCMDYODDPHBG-ACZMJKKPSA-N Gln-Ala-Cys Chemical compound [H]N[C@@H](CCC(N)=O)C(=O)N[C@@H](C)C(=O)N[C@@H](CS)C(O)=O DTCCMDYODDPHBG-ACZMJKKPSA-N 0.000 description 1
- MLZRSFQRBDNJON-GUBZILKMSA-N Gln-Ala-Lys Chemical compound C[C@@H](C(=O)N[C@@H](CCCCN)C(=O)O)NC(=O)[C@H](CCC(=O)N)N MLZRSFQRBDNJON-GUBZILKMSA-N 0.000 description 1
- IGNGBUVODQLMRJ-CIUDSAMLSA-N Gln-Ala-Met Chemical compound [H]N[C@@H](CCC(N)=O)C(=O)N[C@@H](C)C(=O)N[C@@H](CCSC)C(O)=O IGNGBUVODQLMRJ-CIUDSAMLSA-N 0.000 description 1
- KZKBJEUWNMQTLV-XDTLVQLUSA-N Gln-Ala-Tyr Chemical compound [H]N[C@@H](CCC(N)=O)C(=O)N[C@@H](C)C(=O)N[C@@H](CC1=CC=C(O)C=C1)C(O)=O KZKBJEUWNMQTLV-XDTLVQLUSA-N 0.000 description 1
- JSYULGSPLTZDHM-NRPADANISA-N Gln-Ala-Val Chemical compound [H]N[C@@H](CCC(N)=O)C(=O)N[C@@H](C)C(=O)N[C@@H](C(C)C)C(O)=O JSYULGSPLTZDHM-NRPADANISA-N 0.000 description 1
- DLOHWQXXGMEZDW-CIUDSAMLSA-N Gln-Arg-Asn Chemical compound [H]N[C@@H](CCC(N)=O)C(=O)N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CC(N)=O)C(O)=O DLOHWQXXGMEZDW-CIUDSAMLSA-N 0.000 description 1
- PGPJSRSLQNXBDT-YUMQZZPRSA-N Gln-Arg-Gly Chemical compound [H]N[C@@H](CCC(N)=O)C(=O)N[C@@H](CCCNC(N)=N)C(=O)NCC(O)=O PGPJSRSLQNXBDT-YUMQZZPRSA-N 0.000 description 1
- PRBLYKYHAJEABA-SRVKXCTJSA-N Gln-Arg-Leu Chemical compound [H]N[C@@H](CCC(N)=O)C(=O)N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CC(C)C)C(O)=O PRBLYKYHAJEABA-SRVKXCTJSA-N 0.000 description 1
- JFOKLAPFYCTNHW-SRVKXCTJSA-N Gln-Arg-Lys Chemical compound C(CCN)C[C@@H](C(=O)O)NC(=O)[C@H](CCCN=C(N)N)NC(=O)[C@H](CCC(=O)N)N JFOKLAPFYCTNHW-SRVKXCTJSA-N 0.000 description 1
- ZFADFBPRMSBPOT-KKUMJFAQSA-N Gln-Arg-Phe Chemical compound N[C@@H](CCC(N)=O)C(=O)N[C@@H](CCCN=C(N)N)C(=O)N[C@@H](Cc1ccccc1)C(O)=O ZFADFBPRMSBPOT-KKUMJFAQSA-N 0.000 description 1
- JESJDAAGXULQOP-CIUDSAMLSA-N Gln-Arg-Ser Chemical compound C(C[C@@H](C(=O)N[C@@H](CO)C(=O)O)NC(=O)[C@H](CCC(=O)N)N)CN=C(N)N JESJDAAGXULQOP-CIUDSAMLSA-N 0.000 description 1
- WMOMPXKOKASNBK-PEFMBERDSA-N Gln-Asn-Ile Chemical compound [H]N[C@@H](CCC(N)=O)C(=O)N[C@@H](CC(N)=O)C(=O)N[C@@H]([C@@H](C)CC)C(O)=O WMOMPXKOKASNBK-PEFMBERDSA-N 0.000 description 1
- AAOBFSKXAVIORT-GUBZILKMSA-N Gln-Asn-Leu Chemical compound [H]N[C@@H](CCC(N)=O)C(=O)N[C@@H](CC(N)=O)C(=O)N[C@@H](CC(C)C)C(O)=O AAOBFSKXAVIORT-GUBZILKMSA-N 0.000 description 1
- KWLMLNHADZIJIS-CIUDSAMLSA-N Gln-Asn-Met Chemical compound CSCC[C@@H](C(=O)O)NC(=O)[C@H](CC(=O)N)NC(=O)[C@H](CCC(=O)N)N KWLMLNHADZIJIS-CIUDSAMLSA-N 0.000 description 1
- CKNUKHBRCSMKMO-XHNCKOQMSA-N Gln-Asn-Pro Chemical compound C1C[C@@H](N(C1)C(=O)[C@H](CC(=O)N)NC(=O)[C@H](CCC(=O)N)N)C(=O)O CKNUKHBRCSMKMO-XHNCKOQMSA-N 0.000 description 1
- QYTKAVBFRUGYAU-ACZMJKKPSA-N Gln-Asp-Asn Chemical compound [H]N[C@@H](CCC(N)=O)C(=O)N[C@@H](CC(O)=O)C(=O)N[C@@H](CC(N)=O)C(O)=O QYTKAVBFRUGYAU-ACZMJKKPSA-N 0.000 description 1
- CRRFJBGUGNNOCS-PEFMBERDSA-N Gln-Asp-Ile Chemical compound [H]N[C@@H](CCC(N)=O)C(=O)N[C@@H](CC(O)=O)C(=O)N[C@@H]([C@@H](C)CC)C(O)=O CRRFJBGUGNNOCS-PEFMBERDSA-N 0.000 description 1
- KZEUVLLVULIPNX-GUBZILKMSA-N Gln-Asp-Lys Chemical compound C(CCN)C[C@@H](C(=O)O)NC(=O)[C@H](CC(=O)O)NC(=O)[C@H](CCC(=O)N)N KZEUVLLVULIPNX-GUBZILKMSA-N 0.000 description 1
- SOIAHPSKKUYREP-CIUDSAMLSA-N Gln-Asp-Met Chemical compound CSCC[C@@H](C(=O)O)NC(=O)[C@H](CC(=O)O)NC(=O)[C@H](CCC(=O)N)N SOIAHPSKKUYREP-CIUDSAMLSA-N 0.000 description 1
- PCKOTDPDHIBGRW-CIUDSAMLSA-N Gln-Cys-Arg Chemical compound C(C[C@@H](C(=O)O)NC(=O)[C@H](CS)NC(=O)[C@H](CCC(=O)N)N)CN=C(N)N PCKOTDPDHIBGRW-CIUDSAMLSA-N 0.000 description 1
- FJAYYNIXQNERSO-ACZMJKKPSA-N Gln-Cys-Asp Chemical compound C(CC(=O)N)[C@@H](C(=O)N[C@@H](CS)C(=O)N[C@@H](CC(=O)O)C(=O)O)N FJAYYNIXQNERSO-ACZMJKKPSA-N 0.000 description 1
- IPHGBVYWRKCGKG-FXQIFTODSA-N Gln-Cys-Glu Chemical compound NC(=O)CC[C@H](N)C(=O)N[C@@H](CS)C(=O)N[C@@H](CCC(O)=O)C(O)=O IPHGBVYWRKCGKG-FXQIFTODSA-N 0.000 description 1
- UVAOVENCIONMJP-GUBZILKMSA-N Gln-Cys-Leu Chemical compound [H]N[C@@H](CCC(N)=O)C(=O)N[C@@H](CS)C(=O)N[C@@H](CC(C)C)C(O)=O UVAOVENCIONMJP-GUBZILKMSA-N 0.000 description 1
- VNCLJDOTEPPBBD-GUBZILKMSA-N Gln-Cys-Lys Chemical compound C(CCN)C[C@@H](C(=O)O)NC(=O)[C@H](CS)NC(=O)[C@H](CCC(=O)N)N VNCLJDOTEPPBBD-GUBZILKMSA-N 0.000 description 1
- ZDJZEGYVKANKED-NRPADANISA-N Gln-Cys-Val Chemical compound [H]N[C@@H](CCC(N)=O)C(=O)N[C@@H](CS)C(=O)N[C@@H](C(C)C)C(O)=O ZDJZEGYVKANKED-NRPADANISA-N 0.000 description 1
- LOJYQMFIIJVETK-WDSKDSINSA-N Gln-Gln Chemical compound NC(=O)CC[C@H](N)C(=O)N[C@@H](CCC(N)=O)C(O)=O LOJYQMFIIJVETK-WDSKDSINSA-N 0.000 description 1
- LPYPANUXJGFMGV-FXQIFTODSA-N Gln-Gln-Ala Chemical compound C[C@@H](C(=O)O)NC(=O)[C@H](CCC(=O)N)NC(=O)[C@H](CCC(=O)N)N LPYPANUXJGFMGV-FXQIFTODSA-N 0.000 description 1
- KVXVVDFOZNYYKZ-DCAQKATOSA-N Gln-Gln-Leu Chemical compound [H]N[C@@H](CCC(N)=O)C(=O)N[C@@H](CCC(N)=O)C(=O)N[C@@H](CC(C)C)C(O)=O KVXVVDFOZNYYKZ-DCAQKATOSA-N 0.000 description 1
- RBWKVOSARCFSQQ-FXQIFTODSA-N Gln-Gln-Ser Chemical compound NC(=O)CC[C@H](N)C(=O)N[C@@H](CCC(N)=O)C(=O)N[C@@H](CO)C(O)=O RBWKVOSARCFSQQ-FXQIFTODSA-N 0.000 description 1
- UFNSPPFJOHNXRE-AUTRQRHGSA-N Gln-Gln-Val Chemical compound [H]N[C@@H](CCC(N)=O)C(=O)N[C@@H](CCC(N)=O)C(=O)N[C@@H](C(C)C)C(O)=O UFNSPPFJOHNXRE-AUTRQRHGSA-N 0.000 description 1
- ZNZPKVQURDQFFS-FXQIFTODSA-N Gln-Glu-Ser Chemical compound [H]N[C@@H](CCC(N)=O)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CO)C(O)=O ZNZPKVQURDQFFS-FXQIFTODSA-N 0.000 description 1
- WVUZERSNWGUKJY-BPUTZDHNSA-N Gln-Glu-Trp Chemical compound C1=CC=C2C(=C1)C(=CN2)C[C@@H](C(=O)O)NC(=O)[C@H](CCC(=O)O)NC(=O)[C@H](CCC(=O)N)N WVUZERSNWGUKJY-BPUTZDHNSA-N 0.000 description 1
- XJKAKYXMFHUIHT-AUTRQRHGSA-N Gln-Glu-Val Chemical compound CC(C)[C@@H](C(=O)O)NC(=O)[C@H](CCC(=O)O)NC(=O)[C@H](CCC(=O)N)N XJKAKYXMFHUIHT-AUTRQRHGSA-N 0.000 description 1
- VSXBYIJUAXPAAL-WDSKDSINSA-N Gln-Gly-Ala Chemical compound OC(=O)[C@H](C)NC(=O)CNC(=O)[C@@H](N)CCC(N)=O VSXBYIJUAXPAAL-WDSKDSINSA-N 0.000 description 1
- MFJAPSYJQJCQDN-BQBZGAKWSA-N Gln-Gly-Glu Chemical compound NC(=O)CC[C@H](N)C(=O)NCC(=O)N[C@@H](CCC(O)=O)C(O)=O MFJAPSYJQJCQDN-BQBZGAKWSA-N 0.000 description 1
- VGTDBGYFVWOQTI-RYUDHWBXSA-N Gln-Gly-Phe Chemical compound NC(=O)CC[C@H](N)C(=O)NCC(=O)N[C@H](C(O)=O)CC1=CC=CC=C1 VGTDBGYFVWOQTI-RYUDHWBXSA-N 0.000 description 1
- JNEITCMDYWKPIW-GUBZILKMSA-N Gln-His-Cys Chemical compound C1=C(NC=N1)C[C@@H](C(=O)N[C@@H](CS)C(=O)O)NC(=O)[C@H](CCC(=O)N)N JNEITCMDYWKPIW-GUBZILKMSA-N 0.000 description 1
- PODFFOWWLUPNMN-DCAQKATOSA-N Gln-His-Gln Chemical compound [H]N[C@@H](CCC(N)=O)C(=O)N[C@@H](CC1=CNC=N1)C(=O)N[C@@H](CCC(N)=O)C(O)=O PODFFOWWLUPNMN-DCAQKATOSA-N 0.000 description 1
- SBHVGKBYOQKAEA-SDDRHHMPSA-N Gln-His-Pro Chemical compound C1C[C@@H](N(C1)C(=O)[C@H](CC2=CN=CN2)NC(=O)[C@H](CCC(=O)N)N)C(=O)O SBHVGKBYOQKAEA-SDDRHHMPSA-N 0.000 description 1
- TWTWUBHEWQPMQW-ZPFDUUQYSA-N Gln-Ile-Arg Chemical compound [H]N[C@@H](CCC(N)=O)C(=O)N[C@@H]([C@@H](C)CC)C(=O)N[C@@H](CCCNC(N)=N)C(O)=O TWTWUBHEWQPMQW-ZPFDUUQYSA-N 0.000 description 1
- DAAUVRPSZRDMBV-KBIXCLLPSA-N Gln-Ile-Cys Chemical compound CC[C@H](C)[C@@H](C(=O)N[C@@H](CS)C(=O)O)NC(=O)[C@H](CCC(=O)N)N DAAUVRPSZRDMBV-KBIXCLLPSA-N 0.000 description 1
- HXOLDXKNWKLDMM-YVNDNENWSA-N Gln-Ile-Glu Chemical compound CC[C@H](C)[C@@H](C(=O)N[C@@H](CCC(=O)O)C(=O)O)NC(=O)[C@H](CCC(=O)N)N HXOLDXKNWKLDMM-YVNDNENWSA-N 0.000 description 1
- KSKFIECUYMYWNS-AVGNSLFASA-N Gln-Lys-His Chemical compound C1=C(NC=N1)C[C@@H](C(=O)O)NC(=O)[C@H](CCCCN)NC(=O)[C@H](CCC(=O)N)N KSKFIECUYMYWNS-AVGNSLFASA-N 0.000 description 1
- KLKYKPXITJBSNI-CIUDSAMLSA-N Gln-Met-Ala Chemical compound [H]N[C@@H](CCC(N)=O)C(=O)N[C@@H](CCSC)C(=O)N[C@@H](C)C(O)=O KLKYKPXITJBSNI-CIUDSAMLSA-N 0.000 description 1
- LUGUNEGJNDEBLU-DCAQKATOSA-N Gln-Met-Arg Chemical compound CSCC[C@@H](C(=O)N[C@@H](CCCN=C(N)N)C(=O)O)NC(=O)[C@H](CCC(=O)N)N LUGUNEGJNDEBLU-DCAQKATOSA-N 0.000 description 1
- HHRAEXBUNGTOGZ-IHRRRGAJSA-N Gln-Phe-Gln Chemical compound [H]N[C@@H](CCC(N)=O)C(=O)N[C@@H](CC1=CC=CC=C1)C(=O)N[C@@H](CCC(N)=O)C(O)=O HHRAEXBUNGTOGZ-IHRRRGAJSA-N 0.000 description 1
- FTTHLXOMDMLKKW-FHWLQOOXSA-N Gln-Phe-Phe Chemical compound C([C@H](NC(=O)[C@H](CCC(N)=O)N)C(=O)N[C@@H](CC=1C=CC=CC=1)C(O)=O)C1=CC=CC=C1 FTTHLXOMDMLKKW-FHWLQOOXSA-N 0.000 description 1
- FQCILXROGNOZON-YUMQZZPRSA-N Gln-Pro-Gly Chemical compound NC(=O)CC[C@H](N)C(=O)N1CCC[C@H]1C(=O)NCC(O)=O FQCILXROGNOZON-YUMQZZPRSA-N 0.000 description 1
- XQDGOJPVMSWZSO-SRVKXCTJSA-N Gln-Pro-Leu Chemical compound CC(C)C[C@@H](C(=O)O)NC(=O)[C@@H]1CCCN1C(=O)[C@H](CCC(=O)N)N XQDGOJPVMSWZSO-SRVKXCTJSA-N 0.000 description 1
- NYCVMJGIJYQWDO-CIUDSAMLSA-N Gln-Ser-Arg Chemical compound [H]N[C@@H](CCC(N)=O)C(=O)N[C@@H](CO)C(=O)N[C@@H](CCCNC(N)=N)C(O)=O NYCVMJGIJYQWDO-CIUDSAMLSA-N 0.000 description 1
- UTOQQOMEJDPDMX-ACZMJKKPSA-N Gln-Ser-Asp Chemical compound [H]N[C@@H](CCC(N)=O)C(=O)N[C@@H](CO)C(=O)N[C@@H](CC(O)=O)C(O)=O UTOQQOMEJDPDMX-ACZMJKKPSA-N 0.000 description 1
- MFHVAWMMKZBSRQ-ACZMJKKPSA-N Gln-Ser-Cys Chemical compound C(CC(=O)N)[C@@H](C(=O)N[C@@H](CO)C(=O)N[C@@H](CS)C(=O)O)N MFHVAWMMKZBSRQ-ACZMJKKPSA-N 0.000 description 1
- KVQOVQVGVKDZNW-GUBZILKMSA-N Gln-Ser-His Chemical compound C1=C(NC=N1)C[C@@H](C(=O)O)NC(=O)[C@H](CO)NC(=O)[C@H](CCC(=O)N)N KVQOVQVGVKDZNW-GUBZILKMSA-N 0.000 description 1
- SYZZMPFLOLSMHL-XHNCKOQMSA-N Gln-Ser-Pro Chemical compound C1C[C@@H](N(C1)C(=O)[C@H](CO)NC(=O)[C@H](CCC(=O)N)N)C(=O)O SYZZMPFLOLSMHL-XHNCKOQMSA-N 0.000 description 1
- GHAXJVNBAKGWEJ-AVGNSLFASA-N Gln-Ser-Tyr Chemical compound [H]N[C@@H](CCC(N)=O)C(=O)N[C@@H](CO)C(=O)N[C@@H](CC1=CC=C(O)C=C1)C(O)=O GHAXJVNBAKGWEJ-AVGNSLFASA-N 0.000 description 1
- GTBXHETZPUURJE-KKUMJFAQSA-N Gln-Tyr-Arg Chemical compound [H]N[C@@H](CCC(N)=O)C(=O)N[C@@H](CC1=CC=C(O)C=C1)C(=O)N[C@@H](CCCNC(N)=N)C(O)=O GTBXHETZPUURJE-KKUMJFAQSA-N 0.000 description 1
- SGVGIVDZLSHSEN-RYUDHWBXSA-N Gln-Tyr-Gly Chemical compound [H]N[C@@H](CCC(N)=O)C(=O)N[C@@H](CC1=CC=C(O)C=C1)C(=O)NCC(O)=O SGVGIVDZLSHSEN-RYUDHWBXSA-N 0.000 description 1
- VCUNGPMMPNJSGS-JYJNAYRXSA-N Gln-Tyr-Lys Chemical compound C1=CC(=CC=C1C[C@@H](C(=O)N[C@@H](CCCCN)C(=O)O)NC(=O)[C@H](CCC(=O)N)N)O VCUNGPMMPNJSGS-JYJNAYRXSA-N 0.000 description 1
- OACPJRQRAHMQEQ-NHCYSSNCSA-N Gln-Val-Arg Chemical compound NC(=O)CC[C@H](N)C(=O)N[C@@H](C(C)C)C(=O)N[C@@H](CCCN=C(N)N)C(O)=O OACPJRQRAHMQEQ-NHCYSSNCSA-N 0.000 description 1
- BBFCMGBMYIAGRS-AUTRQRHGSA-N Gln-Val-Glu Chemical compound [H]N[C@@H](CCC(N)=O)C(=O)N[C@@H](C(C)C)C(=O)N[C@@H](CCC(O)=O)C(O)=O BBFCMGBMYIAGRS-AUTRQRHGSA-N 0.000 description 1
- MKRDNSWGJWTBKZ-GVXVVHGQSA-N Gln-Val-Lys Chemical compound CC(C)[C@@H](C(=O)N[C@@H](CCCCN)C(=O)O)NC(=O)[C@H](CCC(=O)N)N MKRDNSWGJWTBKZ-GVXVVHGQSA-N 0.000 description 1
- FITIQFSXXBKFFM-NRPADANISA-N Gln-Val-Ser Chemical compound [H]N[C@@H](CCC(N)=O)C(=O)N[C@@H](C(C)C)C(=O)N[C@@H](CO)C(O)=O FITIQFSXXBKFFM-NRPADANISA-N 0.000 description 1
- FHPXTPQBODWBIY-CIUDSAMLSA-N Glu-Ala-Arg Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N[C@@H](C)C(=O)N[C@@H](CCCNC(N)=N)C(O)=O FHPXTPQBODWBIY-CIUDSAMLSA-N 0.000 description 1
- SZXSSXUNOALWCH-ACZMJKKPSA-N Glu-Ala-Asn Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N[C@@H](C)C(=O)N[C@@H](CC(N)=O)C(O)=O SZXSSXUNOALWCH-ACZMJKKPSA-N 0.000 description 1
- WZZSKAJIHTUUSG-ACZMJKKPSA-N Glu-Ala-Asp Chemical compound OC(=O)C[C@@H](C(O)=O)NC(=O)[C@H](C)NC(=O)[C@@H](N)CCC(O)=O WZZSKAJIHTUUSG-ACZMJKKPSA-N 0.000 description 1
- UTKICHUQEQBDGC-ACZMJKKPSA-N Glu-Ala-Cys Chemical compound C[C@@H](C(=O)N[C@@H](CS)C(=O)O)NC(=O)[C@H](CCC(=O)O)N UTKICHUQEQBDGC-ACZMJKKPSA-N 0.000 description 1
- LKDIBBOKUAASNP-FXQIFTODSA-N Glu-Ala-Glu Chemical compound OC(=O)CC[C@H](N)C(=O)N[C@@H](C)C(=O)N[C@@H](CCC(O)=O)C(O)=O LKDIBBOKUAASNP-FXQIFTODSA-N 0.000 description 1
- MXOODARRORARSU-ACZMJKKPSA-N Glu-Ala-Ser Chemical compound C[C@@H](C(=O)N[C@@H](CO)C(=O)O)NC(=O)[C@H](CCC(=O)O)N MXOODARRORARSU-ACZMJKKPSA-N 0.000 description 1
- CVPXINNKRTZBMO-CIUDSAMLSA-N Glu-Arg-Asn Chemical compound C(C[C@@H](C(=O)N[C@@H](CC(=O)N)C(=O)O)NC(=O)[C@H](CCC(=O)O)N)CN=C(N)N CVPXINNKRTZBMO-CIUDSAMLSA-N 0.000 description 1
- WOSRKEJQESVHGA-CIUDSAMLSA-N Glu-Arg-Ser Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CO)C(O)=O WOSRKEJQESVHGA-CIUDSAMLSA-N 0.000 description 1
- SRZLHYPAOXBBSB-HJGDQZAQSA-N Glu-Arg-Thr Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N[C@@H](CCCNC(N)=N)C(=O)N[C@@H]([C@@H](C)O)C(O)=O SRZLHYPAOXBBSB-HJGDQZAQSA-N 0.000 description 1
- AKJRHDMTEJXTPV-ACZMJKKPSA-N Glu-Asn-Ala Chemical compound C[C@H](NC(=O)[C@H](CC(N)=O)NC(=O)[C@@H](N)CCC(O)=O)C(O)=O AKJRHDMTEJXTPV-ACZMJKKPSA-N 0.000 description 1
- GLWXKFRTOHKGIT-ACZMJKKPSA-N Glu-Asn-Asn Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N[C@@H](CC(N)=O)C(=O)N[C@@H](CC(N)=O)C(O)=O GLWXKFRTOHKGIT-ACZMJKKPSA-N 0.000 description 1
- CKRUHITYRFNUKW-WDSKDSINSA-N Glu-Asn-Gly Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N[C@@H](CC(N)=O)C(=O)NCC(O)=O CKRUHITYRFNUKW-WDSKDSINSA-N 0.000 description 1
- RJONUNZIMUXUOI-GUBZILKMSA-N Glu-Asn-Lys Chemical compound C(CCN)C[C@@H](C(=O)O)NC(=O)[C@H](CC(=O)N)NC(=O)[C@H](CCC(=O)O)N RJONUNZIMUXUOI-GUBZILKMSA-N 0.000 description 1
- ZJICFHQSPWFBKP-AVGNSLFASA-N Glu-Asn-Tyr Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N[C@@H](CC(N)=O)C(=O)N[C@@H](CC1=CC=C(O)C=C1)C(O)=O ZJICFHQSPWFBKP-AVGNSLFASA-N 0.000 description 1
- QPRZKNOOOBWXSU-CIUDSAMLSA-N Glu-Asp-Arg Chemical compound OC(=O)CC[C@H](N)C(=O)N[C@@H](CC(O)=O)C(=O)N[C@H](C(O)=O)CCCN=C(N)N QPRZKNOOOBWXSU-CIUDSAMLSA-N 0.000 description 1
- VAIWPXWHWAPYDF-FXQIFTODSA-N Glu-Asp-Gln Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N[C@@H](CC(O)=O)C(=O)N[C@@H](CCC(N)=O)C(O)=O VAIWPXWHWAPYDF-FXQIFTODSA-N 0.000 description 1
- PAQUJCSYVIBPLC-AVGNSLFASA-N Glu-Asp-Phe Chemical compound OC(=O)CC[C@H](N)C(=O)N[C@@H](CC(O)=O)C(=O)N[C@H](C(O)=O)CC1=CC=CC=C1 PAQUJCSYVIBPLC-AVGNSLFASA-N 0.000 description 1
- CKOFNWCLWRYUHK-XHNCKOQMSA-N Glu-Asp-Pro Chemical compound C1C[C@@H](N(C1)C(=O)[C@H](CC(=O)O)NC(=O)[C@H](CCC(=O)O)N)C(=O)O CKOFNWCLWRYUHK-XHNCKOQMSA-N 0.000 description 1
- JRCUFCXYZLPSDZ-ACZMJKKPSA-N Glu-Asp-Ser Chemical compound OC(=O)CC[C@H](N)C(=O)N[C@@H](CC(O)=O)C(=O)N[C@@H](CO)C(O)=O JRCUFCXYZLPSDZ-ACZMJKKPSA-N 0.000 description 1
- FLQAKQOBSPFGKG-CIUDSAMLSA-N Glu-Cys-Arg Chemical compound OC(=O)CC[C@H](N)C(=O)N[C@@H](CS)C(=O)N[C@H](C(O)=O)CCCN=C(N)N FLQAKQOBSPFGKG-CIUDSAMLSA-N 0.000 description 1
- LSTFYPOGBGFIPP-FXQIFTODSA-N Glu-Cys-Gln Chemical compound OC(=O)CC[C@H](N)C(=O)N[C@@H](CS)C(=O)N[C@@H](CCC(N)=O)C(O)=O LSTFYPOGBGFIPP-FXQIFTODSA-N 0.000 description 1
- ZZIFPJZQHRJERU-WDSKDSINSA-N Glu-Cys-Gly Chemical compound OC(=O)CC[C@H](N)C(=O)N[C@@H](CS)C(=O)NCC(O)=O ZZIFPJZQHRJERU-WDSKDSINSA-N 0.000 description 1
- OXEMJGCAJFFREE-FXQIFTODSA-N Glu-Gln-Ala Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N[C@@H](CCC(N)=O)C(=O)N[C@@H](C)C(O)=O OXEMJGCAJFFREE-FXQIFTODSA-N 0.000 description 1
- PVBBEKPHARMPHX-DCAQKATOSA-N Glu-Gln-Leu Chemical compound CC(C)C[C@@H](C(O)=O)NC(=O)[C@H](CCC(N)=O)NC(=O)[C@@H](N)CCC(O)=O PVBBEKPHARMPHX-DCAQKATOSA-N 0.000 description 1
- HTTSBEBKVNEDFE-AUTRQRHGSA-N Glu-Gln-Val Chemical compound CC(C)[C@@H](C(=O)O)NC(=O)[C@H](CCC(=O)N)NC(=O)[C@H](CCC(=O)O)N HTTSBEBKVNEDFE-AUTRQRHGSA-N 0.000 description 1
- CGOHAEBMDSEKFB-FXQIFTODSA-N Glu-Glu-Ala Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](C)C(O)=O CGOHAEBMDSEKFB-FXQIFTODSA-N 0.000 description 1
- MUSGDMDGNGXULI-DCAQKATOSA-N Glu-Glu-Leu Chemical compound CC(C)C[C@@H](C(O)=O)NC(=O)[C@H](CCC(O)=O)NC(=O)[C@@H](N)CCC(O)=O MUSGDMDGNGXULI-DCAQKATOSA-N 0.000 description 1
- OGNJZUXUTPQVBR-BQBZGAKWSA-N Glu-Gly-Glu Chemical compound OC(=O)CC[C@H](N)C(=O)NCC(=O)N[C@@H](CCC(O)=O)C(O)=O OGNJZUXUTPQVBR-BQBZGAKWSA-N 0.000 description 1
- ZWQVYZXPYSYPJD-RYUDHWBXSA-N Glu-Gly-Phe Chemical compound OC(=O)CC[C@H](N)C(=O)NCC(=O)N[C@H](C(O)=O)CC1=CC=CC=C1 ZWQVYZXPYSYPJD-RYUDHWBXSA-N 0.000 description 1
- RAUDKMVXNOWDLS-WDSKDSINSA-N Glu-Gly-Ser Chemical compound OC(=O)CC[C@H](N)C(=O)NCC(=O)N[C@@H](CO)C(O)=O RAUDKMVXNOWDLS-WDSKDSINSA-N 0.000 description 1
- HILMIYALTUQTRC-XVKPBYJWSA-N Glu-Gly-Val Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)NCC(=O)N[C@@H](C(C)C)C(O)=O HILMIYALTUQTRC-XVKPBYJWSA-N 0.000 description 1
- ZJFNRQHUIHKZJF-GUBZILKMSA-N Glu-His-Asp Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N[C@@H](CC1=CNC=N1)C(=O)N[C@@H](CC(O)=O)C(O)=O ZJFNRQHUIHKZJF-GUBZILKMSA-N 0.000 description 1
- JGHNIWVNCAOVRO-DCAQKATOSA-N Glu-His-Glu Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N[C@@H](CC1=CNC=N1)C(=O)N[C@@H](CCC(O)=O)C(O)=O JGHNIWVNCAOVRO-DCAQKATOSA-N 0.000 description 1
- CXRWMMRLEMVSEH-PEFMBERDSA-N Glu-Ile-Asn Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N[C@@H]([C@@H](C)CC)C(=O)N[C@@H](CC(N)=O)C(O)=O CXRWMMRLEMVSEH-PEFMBERDSA-N 0.000 description 1
- XTZDZAXYPDISRR-MNXVOIDGSA-N Glu-Ile-Lys Chemical compound CC[C@H](C)[C@@H](C(=O)N[C@@H](CCCCN)C(=O)O)NC(=O)[C@H](CCC(=O)O)N XTZDZAXYPDISRR-MNXVOIDGSA-N 0.000 description 1
- INGJLBQKTRJLFO-UKJIMTQDSA-N Glu-Ile-Val Chemical compound CC(C)[C@@H](C(O)=O)NC(=O)[C@H]([C@@H](C)CC)NC(=O)[C@@H](N)CCC(O)=O INGJLBQKTRJLFO-UKJIMTQDSA-N 0.000 description 1
- VSRCAOIHMGCIJK-SRVKXCTJSA-N Glu-Leu-Arg Chemical compound OC(=O)CC[C@H](N)C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](CCCN=C(N)N)C(O)=O VSRCAOIHMGCIJK-SRVKXCTJSA-N 0.000 description 1
- NWOUBJNMZDDGDT-AVGNSLFASA-N Glu-Leu-His Chemical compound OC(=O)CC[C@H](N)C(=O)N[C@@H](CC(C)C)C(=O)N[C@H](C(O)=O)CC1=CN=CN1 NWOUBJNMZDDGDT-AVGNSLFASA-N 0.000 description 1
- GJBUAAAIZSRCDC-GVXVVHGQSA-N Glu-Leu-Val Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](C(C)C)C(O)=O GJBUAAAIZSRCDC-GVXVVHGQSA-N 0.000 description 1
- CUPSDFQZTVVTSK-GUBZILKMSA-N Glu-Lys-Asp Chemical compound OC(=O)C[C@@H](C(O)=O)NC(=O)[C@H](CCCCN)NC(=O)[C@@H](N)CCC(O)=O CUPSDFQZTVVTSK-GUBZILKMSA-N 0.000 description 1
- RBXSZQRSEGYDFG-GUBZILKMSA-N Glu-Lys-Ser Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N[C@@H](CCCCN)C(=O)N[C@@H](CO)C(O)=O RBXSZQRSEGYDFG-GUBZILKMSA-N 0.000 description 1
- ZQYZDDXTNQXUJH-CIUDSAMLSA-N Glu-Met-Ala Chemical compound C[C@@H](C(=O)O)NC(=O)[C@H](CCSC)NC(=O)[C@H](CCC(=O)O)N ZQYZDDXTNQXUJH-CIUDSAMLSA-N 0.000 description 1
- UERORLSAFUHDGU-AVGNSLFASA-N Glu-Phe-Asn Chemical compound C1=CC=C(C=C1)C[C@@H](C(=O)N[C@@H](CC(=O)N)C(=O)O)NC(=O)[C@H](CCC(=O)O)N UERORLSAFUHDGU-AVGNSLFASA-N 0.000 description 1
- ARIORLIIMJACKZ-KKUMJFAQSA-N Glu-Pro-Tyr Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N1CCC[C@H]1C(=O)N[C@@H](CC1=CC=C(O)C=C1)C(O)=O ARIORLIIMJACKZ-KKUMJFAQSA-N 0.000 description 1
- WIKMTDVSCUJIPJ-CIUDSAMLSA-N Glu-Ser-Arg Chemical compound OC(=O)CC[C@H](N)C(=O)N[C@@H](CO)C(=O)N[C@H](C(O)=O)CCCN=C(N)N WIKMTDVSCUJIPJ-CIUDSAMLSA-N 0.000 description 1
- MRWYPDWDZSLWJM-ACZMJKKPSA-N Glu-Ser-Asp Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N[C@@H](CO)C(=O)N[C@@H](CC(O)=O)C(O)=O MRWYPDWDZSLWJM-ACZMJKKPSA-N 0.000 description 1
- WXONSNSSBYQGNN-AVGNSLFASA-N Glu-Ser-Tyr Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N[C@@H](CO)C(=O)N[C@@H](CC1=CC=C(O)C=C1)C(O)=O WXONSNSSBYQGNN-AVGNSLFASA-N 0.000 description 1
- QCMVGXDELYMZET-GLLZPBPUSA-N Glu-Thr-Glu Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N[C@@H]([C@@H](C)O)C(=O)N[C@@H](CCC(O)=O)C(O)=O QCMVGXDELYMZET-GLLZPBPUSA-N 0.000 description 1
- DLISPGXMKZTWQG-IFFSRLJSSA-N Glu-Thr-Val Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N[C@@H]([C@@H](C)O)C(=O)N[C@@H](C(C)C)C(O)=O DLISPGXMKZTWQG-IFFSRLJSSA-N 0.000 description 1
- VJVAQZYGLMJPTK-QEJZJMRPSA-N Glu-Trp-Asp Chemical compound C1=CC=C2C(=C1)C(=CN2)C[C@@H](C(=O)N[C@@H](CC(=O)O)C(=O)O)NC(=O)[C@H](CCC(=O)O)N VJVAQZYGLMJPTK-QEJZJMRPSA-N 0.000 description 1
- ZTNHPMZHAILHRB-JSGCOSHPSA-N Glu-Trp-Gly Chemical compound C1=CC=C2C(C[C@H](NC(=O)[C@H](CCC(O)=O)N)C(=O)NCC(O)=O)=CNC2=C1 ZTNHPMZHAILHRB-JSGCOSHPSA-N 0.000 description 1
- UCZXXMREFIETQW-AVGNSLFASA-N Glu-Tyr-Asn Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N[C@@H](CC1=CC=C(O)C=C1)C(=O)N[C@@H](CC(N)=O)C(O)=O UCZXXMREFIETQW-AVGNSLFASA-N 0.000 description 1
- VIPDPMHGICREIS-GVXVVHGQSA-N Glu-Val-Leu Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N[C@@H](C(C)C)C(=O)N[C@@H](CC(C)C)C(O)=O VIPDPMHGICREIS-GVXVVHGQSA-N 0.000 description 1
- WQZGKKKJIJFFOK-GASJEMHNSA-N Glucose Natural products OC[C@H]1OC(O)[C@H](O)[C@@H](O)[C@@H]1O WQZGKKKJIJFFOK-GASJEMHNSA-N 0.000 description 1
- VSVZIEVNUYDAFR-YUMQZZPRSA-N Gly-Ala-Leu Chemical compound CC(C)C[C@@H](C(O)=O)NC(=O)[C@H](C)NC(=O)CN VSVZIEVNUYDAFR-YUMQZZPRSA-N 0.000 description 1
- QSDKBRMVXSWAQE-BFHQHQDPSA-N Gly-Ala-Thr Chemical compound C[C@@H](O)[C@@H](C(O)=O)NC(=O)[C@H](C)NC(=O)CN QSDKBRMVXSWAQE-BFHQHQDPSA-N 0.000 description 1
- PYUCNHJQQVSPGN-BQBZGAKWSA-N Gly-Arg-Cys Chemical compound C(C[C@@H](C(=O)N[C@@H](CS)C(=O)O)NC(=O)CN)CN=C(N)N PYUCNHJQQVSPGN-BQBZGAKWSA-N 0.000 description 1
- JPXNYFOHTHSREU-UWVGGRQHSA-N Gly-Arg-His Chemical compound C1=C(NC=N1)C[C@@H](C(=O)O)NC(=O)[C@H](CCCN=C(N)N)NC(=O)CN JPXNYFOHTHSREU-UWVGGRQHSA-N 0.000 description 1
- OCQUNKSFDYDXBG-QXEWZRGKSA-N Gly-Arg-Ile Chemical compound CC[C@H](C)[C@@H](C(O)=O)NC(=O)[C@@H](NC(=O)CN)CCCN=C(N)N OCQUNKSFDYDXBG-QXEWZRGKSA-N 0.000 description 1
- OVSKVOOUFAKODB-UWVGGRQHSA-N Gly-Arg-Leu Chemical compound CC(C)C[C@@H](C(O)=O)NC(=O)[C@@H](NC(=O)CN)CCCN=C(N)N OVSKVOOUFAKODB-UWVGGRQHSA-N 0.000 description 1
- KKBWDNZXYLGJEY-UHFFFAOYSA-N Gly-Arg-Pro Natural products NCC(=O)NC(CCNC(=N)N)C(=O)N1CCCC1C(=O)O KKBWDNZXYLGJEY-UHFFFAOYSA-N 0.000 description 1
- GWCRIHNSVMOBEQ-BQBZGAKWSA-N Gly-Arg-Ser Chemical compound [H]NCC(=O)N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CO)C(O)=O GWCRIHNSVMOBEQ-BQBZGAKWSA-N 0.000 description 1
- XZRZILPOZBVTDB-GJZGRUSLSA-N Gly-Arg-Trp Chemical compound C1=CC=C2C(C[C@H](NC(=O)[C@H](CCCNC(N)=N)NC(=O)CN)C(O)=O)=CNC2=C1 XZRZILPOZBVTDB-GJZGRUSLSA-N 0.000 description 1
- CIMULJZTTOBOPN-WHFBIAKZSA-N Gly-Asn-Asn Chemical compound NCC(=O)N[C@@H](CC(N)=O)C(=O)N[C@@H](CC(N)=O)C(O)=O CIMULJZTTOBOPN-WHFBIAKZSA-N 0.000 description 1
- NZAFOTBEULLEQB-WDSKDSINSA-N Gly-Asn-Glu Chemical compound C(CC(=O)O)[C@@H](C(=O)O)NC(=O)[C@H](CC(=O)N)NC(=O)CN NZAFOTBEULLEQB-WDSKDSINSA-N 0.000 description 1
- JVWPPCWUDRJGAE-YUMQZZPRSA-N Gly-Asn-Leu Chemical compound [H]NCC(=O)N[C@@H](CC(N)=O)C(=O)N[C@@H](CC(C)C)C(O)=O JVWPPCWUDRJGAE-YUMQZZPRSA-N 0.000 description 1
- OCDLPQDYTJPWNG-YUMQZZPRSA-N Gly-Asn-Lys Chemical compound C(CCN)C[C@@H](C(=O)O)NC(=O)[C@H](CC(=O)N)NC(=O)CN OCDLPQDYTJPWNG-YUMQZZPRSA-N 0.000 description 1
- JVACNFOPSUPDTK-QWRGUYRKSA-N Gly-Asn-Phe Chemical compound NCC(=O)N[C@@H](CC(N)=O)C(=O)N[C@H](C(O)=O)CC1=CC=CC=C1 JVACNFOPSUPDTK-QWRGUYRKSA-N 0.000 description 1
- FUTAPPOITCCWTH-WHFBIAKZSA-N Gly-Asp-Asp Chemical compound [H]NCC(=O)N[C@@H](CC(O)=O)C(=O)N[C@@H](CC(O)=O)C(O)=O FUTAPPOITCCWTH-WHFBIAKZSA-N 0.000 description 1
- QSTLUOIOYLYLLF-WDSKDSINSA-N Gly-Asp-Glu Chemical compound [H]NCC(=O)N[C@@H](CC(O)=O)C(=O)N[C@@H](CCC(O)=O)C(O)=O QSTLUOIOYLYLLF-WDSKDSINSA-N 0.000 description 1
- FZQLXNIMCPJVJE-YUMQZZPRSA-N Gly-Asp-Leu Chemical compound [H]NCC(=O)N[C@@H](CC(O)=O)C(=O)N[C@@H](CC(C)C)C(O)=O FZQLXNIMCPJVJE-YUMQZZPRSA-N 0.000 description 1
- LCNXZQROPKFGQK-WHFBIAKZSA-N Gly-Asp-Ser Chemical compound NCC(=O)N[C@@H](CC(O)=O)C(=O)N[C@@H](CO)C(O)=O LCNXZQROPKFGQK-WHFBIAKZSA-N 0.000 description 1
- DTRUBYPMMVPQPD-YUMQZZPRSA-N Gly-Gln-Arg Chemical compound [H]NCC(=O)N[C@@H](CCC(N)=O)C(=O)N[C@@H](CCCNC(N)=N)C(O)=O DTRUBYPMMVPQPD-YUMQZZPRSA-N 0.000 description 1
- KTSZUNRRYXPZTK-BQBZGAKWSA-N Gly-Gln-Glu Chemical compound NCC(=O)N[C@@H](CCC(N)=O)C(=O)N[C@@H](CCC(O)=O)C(O)=O KTSZUNRRYXPZTK-BQBZGAKWSA-N 0.000 description 1
- BYYNJRSNDARRBX-YFKPBYRVSA-N Gly-Gln-Gly Chemical compound NCC(=O)N[C@@H](CCC(N)=O)C(=O)NCC(O)=O BYYNJRSNDARRBX-YFKPBYRVSA-N 0.000 description 1
- AQLHORCVPGXDJW-IUCAKERBSA-N Gly-Gln-Lys Chemical compound C(CCN)C[C@@H](C(=O)O)NC(=O)[C@H](CCC(=O)N)NC(=O)CN AQLHORCVPGXDJW-IUCAKERBSA-N 0.000 description 1
- PABFFPWEJMEVEC-JGVFFNPUSA-N Gly-Gln-Pro Chemical compound C1C[C@@H](N(C1)C(=O)[C@H](CCC(=O)N)NC(=O)CN)C(=O)O PABFFPWEJMEVEC-JGVFFNPUSA-N 0.000 description 1
- SOEATRRYCIPEHA-BQBZGAKWSA-N Gly-Glu-Glu Chemical compound [H]NCC(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CCC(O)=O)C(O)=O SOEATRRYCIPEHA-BQBZGAKWSA-N 0.000 description 1
- XTQFHTHIAKKCTM-YFKPBYRVSA-N Gly-Glu-Gly Chemical compound NCC(=O)N[C@@H](CCC(O)=O)C(=O)NCC(O)=O XTQFHTHIAKKCTM-YFKPBYRVSA-N 0.000 description 1
- JUBDONGMHASUCN-IUCAKERBSA-N Gly-Glu-His Chemical compound NCC(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](Cc1cnc[nH]1)C(O)=O JUBDONGMHASUCN-IUCAKERBSA-N 0.000 description 1
- STVHDEHTKFXBJQ-LAEOZQHASA-N Gly-Glu-Ile Chemical compound [H]NCC(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H]([C@@H](C)CC)C(O)=O STVHDEHTKFXBJQ-LAEOZQHASA-N 0.000 description 1
- LHRXAHLCRMQBGJ-RYUDHWBXSA-N Gly-Glu-Phe Chemical compound C1=CC=C(C=C1)C[C@@H](C(=O)O)NC(=O)[C@H](CCC(=O)O)NC(=O)CN LHRXAHLCRMQBGJ-RYUDHWBXSA-N 0.000 description 1
- QSVCIFZPGLOZGH-WDSKDSINSA-N Gly-Glu-Ser Chemical compound NCC(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CO)C(O)=O QSVCIFZPGLOZGH-WDSKDSINSA-N 0.000 description 1
- GDOZQTNZPCUARW-YFKPBYRVSA-N Gly-Gly-Glu Chemical compound NCC(=O)NCC(=O)N[C@H](C(O)=O)CCC(O)=O GDOZQTNZPCUARW-YFKPBYRVSA-N 0.000 description 1
- PDAWDNVHMUKWJR-ZETCQYMHSA-N Gly-Gly-His Chemical compound NCC(=O)NCC(=O)N[C@H](C(O)=O)CC1=CNC=N1 PDAWDNVHMUKWJR-ZETCQYMHSA-N 0.000 description 1
- SWQALSGKVLYKDT-UHFFFAOYSA-N Gly-Ile-Ala Natural products NCC(=O)NC(C(C)CC)C(=O)NC(C)C(O)=O SWQALSGKVLYKDT-UHFFFAOYSA-N 0.000 description 1
- SXJHOPPTOJACOA-QXEWZRGKSA-N Gly-Ile-Arg Chemical compound NCC(=O)N[C@@H]([C@@H](C)CC)C(=O)N[C@H](C(O)=O)CCCN=C(N)N SXJHOPPTOJACOA-QXEWZRGKSA-N 0.000 description 1
- DGKBSGNCMCLDSL-BYULHYEWSA-N Gly-Ile-Asn Chemical compound CC[C@H](C)[C@@H](C(=O)N[C@@H](CC(=O)N)C(=O)O)NC(=O)CN DGKBSGNCMCLDSL-BYULHYEWSA-N 0.000 description 1
- AAHSHTLISQUZJL-QSFUFRPTSA-N Gly-Ile-Ile Chemical compound [H]NCC(=O)N[C@@H]([C@@H](C)CC)C(=O)N[C@@H]([C@@H](C)CC)C(O)=O AAHSHTLISQUZJL-QSFUFRPTSA-N 0.000 description 1
- UHPAZODVFFYEEL-QWRGUYRKSA-N Gly-Leu-Leu Chemical compound CC(C)C[C@@H](C(O)=O)NC(=O)[C@H](CC(C)C)NC(=O)CN UHPAZODVFFYEEL-QWRGUYRKSA-N 0.000 description 1
- CLNSYANKYVMZNM-UWVGGRQHSA-N Gly-Lys-Arg Chemical compound NCCCC[C@H](NC(=O)CN)C(=O)N[C@H](C(O)=O)CCCN=C(N)N CLNSYANKYVMZNM-UWVGGRQHSA-N 0.000 description 1
- PDUHNKAFQXQNLH-ZETCQYMHSA-N Gly-Lys-Gly Chemical compound NCCCC[C@H](NC(=O)CN)C(=O)NCC(O)=O PDUHNKAFQXQNLH-ZETCQYMHSA-N 0.000 description 1
- MHXKHKWHPNETGG-QWRGUYRKSA-N Gly-Lys-Leu Chemical compound [H]NCC(=O)N[C@@H](CCCCN)C(=O)N[C@@H](CC(C)C)C(O)=O MHXKHKWHPNETGG-QWRGUYRKSA-N 0.000 description 1
- MHZXESQPPXOING-KBPBESRZSA-N Gly-Lys-Phe Chemical compound [H]NCC(=O)N[C@@H](CCCCN)C(=O)N[C@@H](CC1=CC=CC=C1)C(O)=O MHZXESQPPXOING-KBPBESRZSA-N 0.000 description 1
- QGDOOCIPHSSADO-STQMWFEESA-N Gly-Met-Phe Chemical compound [H]NCC(=O)N[C@@H](CCSC)C(=O)N[C@@H](CC1=CC=CC=C1)C(O)=O QGDOOCIPHSSADO-STQMWFEESA-N 0.000 description 1
- IGOYNRWLWHWAQO-JTQLQIEISA-N Gly-Phe-Gly Chemical compound OC(=O)CNC(=O)[C@@H](NC(=O)CN)CC1=CC=CC=C1 IGOYNRWLWHWAQO-JTQLQIEISA-N 0.000 description 1
- IBYOLNARKHMLBG-WHOFXGATSA-N Gly-Phe-Ile Chemical compound CC[C@H](C)[C@@H](C(O)=O)NC(=O)[C@@H](NC(=O)CN)CC1=CC=CC=C1 IBYOLNARKHMLBG-WHOFXGATSA-N 0.000 description 1
- VDCRBJACQKOSMS-JSGCOSHPSA-N Gly-Phe-Val Chemical compound [H]NCC(=O)N[C@@H](CC1=CC=CC=C1)C(=O)N[C@@H](C(C)C)C(O)=O VDCRBJACQKOSMS-JSGCOSHPSA-N 0.000 description 1
- GGLIDLCEPDHEJO-BQBZGAKWSA-N Gly-Pro-Ala Chemical compound OC(=O)[C@H](C)NC(=O)[C@@H]1CCCN1C(=O)CN GGLIDLCEPDHEJO-BQBZGAKWSA-N 0.000 description 1
- JYPCXBJRLBHWME-IUCAKERBSA-N Gly-Pro-Arg Chemical compound NCC(=O)N1CCC[C@H]1C(=O)N[C@@H](CCCNC(N)=N)C(O)=O JYPCXBJRLBHWME-IUCAKERBSA-N 0.000 description 1
- IRJWAYCXIYUHQE-WHFBIAKZSA-N Gly-Ser-Ala Chemical compound OC(=O)[C@H](C)NC(=O)[C@H](CO)NC(=O)CN IRJWAYCXIYUHQE-WHFBIAKZSA-N 0.000 description 1
- ABPRMMYHROQBLY-NKWVEPMBSA-N Gly-Ser-Pro Chemical compound C1C[C@@H](N(C1)C(=O)[C@H](CO)NC(=O)CN)C(=O)O ABPRMMYHROQBLY-NKWVEPMBSA-N 0.000 description 1
- WCORRBXVISTKQL-WHFBIAKZSA-N Gly-Ser-Ser Chemical compound NCC(=O)N[C@@H](CO)C(=O)N[C@@H](CO)C(O)=O WCORRBXVISTKQL-WHFBIAKZSA-N 0.000 description 1
- LCRDMSSAKLTKBU-ZDLURKLDSA-N Gly-Ser-Thr Chemical compound C[C@@H](O)[C@@H](C(O)=O)NC(=O)[C@H](CO)NC(=O)CN LCRDMSSAKLTKBU-ZDLURKLDSA-N 0.000 description 1
- FFJQHWKSGAWSTJ-BFHQHQDPSA-N Gly-Thr-Ala Chemical compound [H]NCC(=O)N[C@@H]([C@@H](C)O)C(=O)N[C@@H](C)C(O)=O FFJQHWKSGAWSTJ-BFHQHQDPSA-N 0.000 description 1
- CUVBTVWFVIIDOC-YEPSODPASA-N Gly-Thr-Val Chemical compound CC(C)[C@@H](C(O)=O)NC(=O)[C@H]([C@@H](C)O)NC(=O)CN CUVBTVWFVIIDOC-YEPSODPASA-N 0.000 description 1
- UIQGJYUEQDOODF-KWQFWETISA-N Gly-Tyr-Ala Chemical compound OC(=O)[C@H](C)NC(=O)[C@@H](NC(=O)CN)CC1=CC=C(O)C=C1 UIQGJYUEQDOODF-KWQFWETISA-N 0.000 description 1
- DUAWRXXTOQOECJ-JSGCOSHPSA-N Gly-Tyr-Val Chemical compound [H]NCC(=O)N[C@@H](CC1=CC=C(O)C=C1)C(=O)N[C@@H](C(C)C)C(O)=O DUAWRXXTOQOECJ-JSGCOSHPSA-N 0.000 description 1
- SBVMXEZQJVUARN-XPUUQOCRSA-N Gly-Val-Ser Chemical compound NCC(=O)N[C@@H](C(C)C)C(=O)N[C@@H](CO)C(O)=O SBVMXEZQJVUARN-XPUUQOCRSA-N 0.000 description 1
- 241000451105 Hakea gibbosa Species 0.000 description 1
- AWHJQEYGWRKPHE-LSJOCFKGSA-N His-Ala-Arg Chemical compound [H]N[C@@H](CC1=CNC=N1)C(=O)N[C@@H](C)C(=O)N[C@@H](CCCNC(N)=N)C(O)=O AWHJQEYGWRKPHE-LSJOCFKGSA-N 0.000 description 1
- VSLXGYMEHVAJBH-DLOVCJGASA-N His-Ala-Leu Chemical compound [H]N[C@@H](CC1=CNC=N1)C(=O)N[C@@H](C)C(=O)N[C@@H](CC(C)C)C(O)=O VSLXGYMEHVAJBH-DLOVCJGASA-N 0.000 description 1
- VCDNHBNNPCDBKV-DLOVCJGASA-N His-Ala-Lys Chemical compound C[C@@H](C(=O)N[C@@H](CCCCN)C(=O)O)NC(=O)[C@H](CC1=CN=CN1)N VCDNHBNNPCDBKV-DLOVCJGASA-N 0.000 description 1
- HXKZJLWGSWQKEA-LSJOCFKGSA-N His-Ala-Val Chemical compound CC(C)[C@@H](C(O)=O)NC(=O)[C@H](C)NC(=O)[C@@H](N)CC1=CN=CN1 HXKZJLWGSWQKEA-LSJOCFKGSA-N 0.000 description 1
- FPNWKONEZAVQJF-GUBZILKMSA-N His-Asn-Gln Chemical compound C1=C(NC=N1)C[C@@H](C(=O)N[C@@H](CC(=O)N)C(=O)N[C@@H](CCC(=O)N)C(=O)O)N FPNWKONEZAVQJF-GUBZILKMSA-N 0.000 description 1
- WMKXFMUJRCEGRP-SRVKXCTJSA-N His-Asn-His Chemical compound C1=C(NC=N1)C[C@@H](C(=O)N[C@@H](CC(=O)N)C(=O)N[C@@H](CC2=CN=CN2)C(=O)O)N WMKXFMUJRCEGRP-SRVKXCTJSA-N 0.000 description 1
- WZOGEMJIZBNFBK-CIUDSAMLSA-N His-Asp-Asn Chemical compound [H]N[C@@H](CC1=CNC=N1)C(=O)N[C@@H](CC(O)=O)C(=O)N[C@@H](CC(N)=O)C(O)=O WZOGEMJIZBNFBK-CIUDSAMLSA-N 0.000 description 1
- UOAVQQRILDGZEN-SRVKXCTJSA-N His-Asp-Leu Chemical compound [H]N[C@@H](CC1=CNC=N1)C(=O)N[C@@H](CC(O)=O)C(=O)N[C@@H](CC(C)C)C(O)=O UOAVQQRILDGZEN-SRVKXCTJSA-N 0.000 description 1
- LDTJBEOANMQRJE-CIUDSAMLSA-N His-Cys-Asp Chemical compound C1=C(NC=N1)C[C@@H](C(=O)N[C@@H](CS)C(=O)N[C@@H](CC(=O)O)C(=O)O)N LDTJBEOANMQRJE-CIUDSAMLSA-N 0.000 description 1
- IDQKGZWUPVOGPZ-GUBZILKMSA-N His-Cys-Gln Chemical compound C1=C(NC=N1)C[C@@H](C(=O)N[C@@H](CS)C(=O)N[C@@H](CCC(=O)N)C(=O)O)N IDQKGZWUPVOGPZ-GUBZILKMSA-N 0.000 description 1
- BQFGKVYHKCNEMF-DCAQKATOSA-N His-Glu-Gln Chemical compound NC(=O)CC[C@@H](C(O)=O)NC(=O)[C@H](CCC(O)=O)NC(=O)[C@@H](N)CC1=CN=CN1 BQFGKVYHKCNEMF-DCAQKATOSA-N 0.000 description 1
- TXLQHACKRLWYCM-DCAQKATOSA-N His-Glu-Glu Chemical compound [H]N[C@@H](CC1=CNC=N1)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CCC(O)=O)C(O)=O TXLQHACKRLWYCM-DCAQKATOSA-N 0.000 description 1
- CHZRWFUGWRTUOD-IUCAKERBSA-N His-Gly-Gln Chemical compound C1=C(NC=N1)C[C@@H](C(=O)NCC(=O)N[C@@H](CCC(=O)N)C(=O)O)N CHZRWFUGWRTUOD-IUCAKERBSA-N 0.000 description 1
- ZUPVLBAXUUGKKN-VHSXEESVSA-N His-Gly-Pro Chemical compound C1C[C@@H](N(C1)C(=O)CNC(=O)[C@H](CC2=CN=CN2)N)C(=O)O ZUPVLBAXUUGKKN-VHSXEESVSA-N 0.000 description 1
- STOOMQFEJUVAKR-KKUMJFAQSA-N His-His-His Chemical compound C([C@H](N)C(=O)N[C@@H](CC=1N=CNC=1)C(=O)N[C@@H](CC=1N=CNC=1)C(O)=O)C1=CNC=N1 STOOMQFEJUVAKR-KKUMJFAQSA-N 0.000 description 1
- 108010093488 His-His-His-His-His-His Proteins 0.000 description 1
- BZKDJRSZWLPJNI-SRVKXCTJSA-N His-His-Ser Chemical compound [H]N[C@@H](CC1=CNC=N1)C(=O)N[C@@H](CC1=CNC=N1)C(=O)N[C@@H](CO)C(O)=O BZKDJRSZWLPJNI-SRVKXCTJSA-N 0.000 description 1
- AIPUZFXMXAHZKY-QWRGUYRKSA-N His-Leu-Gly Chemical compound [H]N[C@@H](CC1=CNC=N1)C(=O)N[C@@H](CC(C)C)C(=O)NCC(O)=O AIPUZFXMXAHZKY-QWRGUYRKSA-N 0.000 description 1
- SKOKHBGDXGTDDP-MELADBBJSA-N His-Leu-Pro Chemical compound CC(C)C[C@@H](C(=O)N1CCC[C@@H]1C(=O)O)NC(=O)[C@H](CC2=CN=CN2)N SKOKHBGDXGTDDP-MELADBBJSA-N 0.000 description 1
- GUXQAPACZVVOKX-AVGNSLFASA-N His-Lys-Gln Chemical compound C1=C(NC=N1)C[C@@H](C(=O)N[C@@H](CCCCN)C(=O)N[C@@H](CCC(=O)N)C(=O)O)N GUXQAPACZVVOKX-AVGNSLFASA-N 0.000 description 1
- NKRWVZQTPXPNRZ-SRVKXCTJSA-N His-Met-Gln Chemical compound NC(=O)CC[C@@H](C(O)=O)NC(=O)[C@H](CCSC)NC(=O)[C@@H](N)CC1=CN=CN1 NKRWVZQTPXPNRZ-SRVKXCTJSA-N 0.000 description 1
- VUUFXXGKMPLKNH-BZSNNMDCSA-N His-Phe-His Chemical compound C1=CC=C(C=C1)C[C@@H](C(=O)N[C@@H](CC2=CN=CN2)C(=O)O)NC(=O)[C@H](CC3=CN=CN3)N VUUFXXGKMPLKNH-BZSNNMDCSA-N 0.000 description 1
- FLXCRBXJRJSDHX-AVGNSLFASA-N His-Pro-Val Chemical compound [H]N[C@@H](CC1=CNC=N1)C(=O)N1CCC[C@H]1C(=O)N[C@@H](C(C)C)C(O)=O FLXCRBXJRJSDHX-AVGNSLFASA-N 0.000 description 1
- PZAJPILZRFPYJJ-SRVKXCTJSA-N His-Ser-Leu Chemical compound [H]N[C@@H](CC1=CNC=N1)C(=O)N[C@@H](CO)C(=O)N[C@@H](CC(C)C)C(O)=O PZAJPILZRFPYJJ-SRVKXCTJSA-N 0.000 description 1
- PBJOQLUVSGXRSW-YTQUADARSA-N His-Trp-Pro Chemical compound C1C[C@@H](N(C1)C(=O)[C@H](CC2=CNC3=CC=CC=C32)NC(=O)[C@H](CC4=CN=CN4)N)C(=O)O PBJOQLUVSGXRSW-YTQUADARSA-N 0.000 description 1
- WSXNWASHQNSMRX-GVXVVHGQSA-N His-Val-Gln Chemical compound CC(C)[C@@H](C(=O)N[C@@H](CCC(=O)N)C(=O)O)NC(=O)[C@H](CC1=CN=CN1)N WSXNWASHQNSMRX-GVXVVHGQSA-N 0.000 description 1
- XGBVLRJLHUVCNK-DCAQKATOSA-N His-Val-Ser Chemical compound [H]N[C@@H](CC1=CNC=N1)C(=O)N[C@@H](C(C)C)C(=O)N[C@@H](CO)C(O)=O XGBVLRJLHUVCNK-DCAQKATOSA-N 0.000 description 1
- 241000282412 Homo Species 0.000 description 1
- 206010020649 Hyperkeratosis Diseases 0.000 description 1
- LQSBBHNVAVNZSX-GHCJXIJMSA-N Ile-Ala-Asn Chemical compound CC[C@H](C)[C@@H](C(=O)N[C@@H](C)C(=O)N[C@@H](CC(=O)N)C(=O)O)N LQSBBHNVAVNZSX-GHCJXIJMSA-N 0.000 description 1
- MKWSZEHGHSLNPF-NAKRPEOUSA-N Ile-Ala-Val Chemical compound CC[C@H](C)[C@@H](C(=O)N[C@@H](C)C(=O)N[C@@H](C(C)C)C(=O)O)N MKWSZEHGHSLNPF-NAKRPEOUSA-N 0.000 description 1
- SACHLUOUHCVIKI-GMOBBJLQSA-N Ile-Arg-Asp Chemical compound CC[C@H](C)[C@@H](C(=O)N[C@@H](CCCN=C(N)N)C(=O)N[C@@H](CC(=O)O)C(=O)O)N SACHLUOUHCVIKI-GMOBBJLQSA-N 0.000 description 1
- ASCFJMSGKUIRDU-ZPFDUUQYSA-N Ile-Arg-Gln Chemical compound CC[C@H](C)[C@H](N)C(=O)N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CCC(N)=O)C(O)=O ASCFJMSGKUIRDU-ZPFDUUQYSA-N 0.000 description 1
- DMHGKBGOUAJRHU-RVMXOQNASA-N Ile-Arg-Pro Chemical compound CC[C@H](C)[C@@H](C(=O)N[C@@H](CCCN=C(N)N)C(=O)N1CCC[C@@H]1C(=O)O)N DMHGKBGOUAJRHU-RVMXOQNASA-N 0.000 description 1
- DMHGKBGOUAJRHU-UHFFFAOYSA-N Ile-Arg-Pro Natural products CCC(C)C(N)C(=O)NC(CCCN=C(N)N)C(=O)N1CCCC1C(O)=O DMHGKBGOUAJRHU-UHFFFAOYSA-N 0.000 description 1
- YKRIXHPEIZUDDY-GMOBBJLQSA-N Ile-Asn-Arg Chemical compound CC[C@H](C)[C@H](N)C(=O)N[C@@H](CC(N)=O)C(=O)N[C@H](C(O)=O)CCCN=C(N)N YKRIXHPEIZUDDY-GMOBBJLQSA-N 0.000 description 1
- IIXDMJNYALIKGP-DJFWLOJKSA-N Ile-Asn-His Chemical compound CC[C@H](C)[C@@H](C(=O)N[C@@H](CC(=O)N)C(=O)N[C@@H](CC1=CN=CN1)C(=O)O)N IIXDMJNYALIKGP-DJFWLOJKSA-N 0.000 description 1
- FJWYJQRCVNGEAQ-ZPFDUUQYSA-N Ile-Asn-Lys Chemical compound CC[C@H](C)[C@@H](C(=O)N[C@@H](CC(=O)N)C(=O)N[C@@H](CCCCN)C(=O)O)N FJWYJQRCVNGEAQ-ZPFDUUQYSA-N 0.000 description 1
- NBJAAWYRLGCJOF-UGYAYLCHSA-N Ile-Asp-Asn Chemical compound CC[C@H](C)[C@@H](C(=O)N[C@@H](CC(=O)O)C(=O)N[C@@H](CC(=O)N)C(=O)O)N NBJAAWYRLGCJOF-UGYAYLCHSA-N 0.000 description 1
- REJKOQYVFDEZHA-SLBDDTMCSA-N Ile-Asp-Trp Chemical compound CC[C@H](C)[C@@H](C(=O)N[C@@H](CC(=O)O)C(=O)N[C@@H](CC1=CNC2=CC=CC=C21)C(=O)O)N REJKOQYVFDEZHA-SLBDDTMCSA-N 0.000 description 1
- FHCNLXMTQJNJNH-KBIXCLLPSA-N Ile-Cys-Gln Chemical compound N[C@@H]([C@@H](C)CC)C(=O)N[C@@H](CS)C(=O)N[C@@H](CCC(N)=O)C(=O)O FHCNLXMTQJNJNH-KBIXCLLPSA-N 0.000 description 1
- VCYVLFAWCJRXFT-HJPIBITLSA-N Ile-Cys-Tyr Chemical compound CC[C@H](C)[C@@H](C(=O)N[C@@H](CS)C(=O)N[C@@H](CC1=CC=C(C=C1)O)C(=O)O)N VCYVLFAWCJRXFT-HJPIBITLSA-N 0.000 description 1
- MTFVYKQRLXYAQN-LAEOZQHASA-N Ile-Glu-Gly Chemical compound [H]N[C@@H]([C@@H](C)CC)C(=O)N[C@@H](CCC(O)=O)C(=O)NCC(O)=O MTFVYKQRLXYAQN-LAEOZQHASA-N 0.000 description 1
- XLCZWMJPVGRWHJ-KQXIARHKSA-N Ile-Glu-Pro Chemical compound CC[C@H](C)[C@@H](C(=O)N[C@@H](CCC(=O)O)C(=O)N1CCC[C@@H]1C(=O)O)N XLCZWMJPVGRWHJ-KQXIARHKSA-N 0.000 description 1
- RIVKTKFVWXRNSJ-GRLWGSQLSA-N Ile-Ile-Gln Chemical compound CC[C@H](C)[C@@H](C(=O)N[C@@H]([C@@H](C)CC)C(=O)N[C@@H](CCC(=O)N)C(=O)O)N RIVKTKFVWXRNSJ-GRLWGSQLSA-N 0.000 description 1
- FZWVCYCYWCLQDH-NHCYSSNCSA-N Ile-Leu-Gly Chemical compound CC[C@H](C)[C@@H](C(=O)N[C@@H](CC(C)C)C(=O)NCC(=O)O)N FZWVCYCYWCLQDH-NHCYSSNCSA-N 0.000 description 1
- TVYWVSJGSHQWMT-AJNGGQMLSA-N Ile-Leu-Lys Chemical compound CC[C@H](C)[C@@H](C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](CCCCN)C(=O)O)N TVYWVSJGSHQWMT-AJNGGQMLSA-N 0.000 description 1
- DSDPLOODKXISDT-XUXIUFHCSA-N Ile-Leu-Val Chemical compound CC[C@H](C)[C@H](N)C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](C(C)C)C(O)=O DSDPLOODKXISDT-XUXIUFHCSA-N 0.000 description 1
- IDMNOFVUXYYZPF-DKIMLUQUSA-N Ile-Lys-Phe Chemical compound CC[C@H](C)[C@@H](C(=O)N[C@@H](CCCCN)C(=O)N[C@@H](CC1=CC=CC=C1)C(=O)O)N IDMNOFVUXYYZPF-DKIMLUQUSA-N 0.000 description 1
- NNVXABCGXOLIEB-PYJNHQTQSA-N Ile-Met-His Chemical compound CC[C@H](C)[C@H](N)C(=O)N[C@@H](CCSC)C(=O)N[C@H](C(O)=O)CC1=CN=CN1 NNVXABCGXOLIEB-PYJNHQTQSA-N 0.000 description 1
- OTSVBELRDMSPKY-PCBIJLKTSA-N Ile-Phe-Asn Chemical compound CC[C@H](C)[C@@H](C(=O)N[C@@H](CC1=CC=CC=C1)C(=O)N[C@@H](CC(=O)N)C(=O)O)N OTSVBELRDMSPKY-PCBIJLKTSA-N 0.000 description 1
- CAHCWMVNBZJVAW-NAKRPEOUSA-N Ile-Pro-Ser Chemical compound CC[C@H](C)[C@@H](C(=O)N1CCC[C@H]1C(=O)N[C@@H](CO)C(=O)O)N CAHCWMVNBZJVAW-NAKRPEOUSA-N 0.000 description 1
- XMYURPUVJSKTMC-KBIXCLLPSA-N Ile-Ser-Gln Chemical compound CC[C@H](C)[C@@H](C(=O)N[C@@H](CO)C(=O)N[C@@H](CCC(=O)N)C(=O)O)N XMYURPUVJSKTMC-KBIXCLLPSA-N 0.000 description 1
- QQVXERGIFIRCGW-NAKRPEOUSA-N Ile-Ser-Met Chemical compound CC[C@H](C)[C@@H](C(=O)N[C@@H](CO)C(=O)N[C@@H](CCSC)C(=O)O)N QQVXERGIFIRCGW-NAKRPEOUSA-N 0.000 description 1
- WXLYNEHOGRYNFU-URLPEUOOSA-N Ile-Thr-Phe Chemical compound CC[C@H](C)[C@@H](C(=O)N[C@@H]([C@@H](C)O)C(=O)N[C@@H](CC1=CC=CC=C1)C(=O)O)N WXLYNEHOGRYNFU-URLPEUOOSA-N 0.000 description 1
- ANTFEOSJMAUGIB-KNZXXDILSA-N Ile-Thr-Pro Chemical compound CC[C@H](C)[C@@H](C(=O)N[C@@H]([C@@H](C)O)C(=O)N1CCC[C@@H]1C(=O)O)N ANTFEOSJMAUGIB-KNZXXDILSA-N 0.000 description 1
- DZMWFIRHFFVBHS-ZEWNOJEFSA-N Ile-Tyr-Phe Chemical compound CC[C@H](C)[C@@H](C(=O)N[C@@H](CC1=CC=C(C=C1)O)C(=O)N[C@@H](CC2=CC=CC=C2)C(=O)O)N DZMWFIRHFFVBHS-ZEWNOJEFSA-N 0.000 description 1
- BCISUQVFDGYZBO-QSFUFRPTSA-N Ile-Val-Asp Chemical compound CC[C@H](C)[C@H](N)C(=O)N[C@@H](C(C)C)C(=O)N[C@H](C(O)=O)CC(O)=O BCISUQVFDGYZBO-QSFUFRPTSA-N 0.000 description 1
- NUEHSWNAFIEBCQ-NAKRPEOUSA-N Ile-Val-Cys Chemical compound CC[C@H](C)[C@@H](C(=O)N[C@@H](C(C)C)C(=O)N[C@@H](CS)C(=O)O)N NUEHSWNAFIEBCQ-NAKRPEOUSA-N 0.000 description 1
- ONIBWKKTOPOVIA-BYPYZUCNSA-N L-Proline Chemical compound OC(=O)[C@@H]1CCCN1 ONIBWKKTOPOVIA-BYPYZUCNSA-N 0.000 description 1
- DCXYFEDJOCDNAF-REOHCLBHSA-N L-asparagine Chemical compound OC(=O)[C@@H](N)CC(N)=O DCXYFEDJOCDNAF-REOHCLBHSA-N 0.000 description 1
- CKLJMWTZIZZHCS-REOHCLBHSA-N L-aspartic acid Chemical compound OC(=O)[C@@H](N)CC(O)=O CKLJMWTZIZZHCS-REOHCLBHSA-N 0.000 description 1
- SENJXOPIZNYLHU-UHFFFAOYSA-N L-leucyl-L-arginine Natural products CC(C)CC(N)C(=O)NC(C(O)=O)CCCN=C(N)N SENJXOPIZNYLHU-UHFFFAOYSA-N 0.000 description 1
- FFEARJCKVFRZRR-BYPYZUCNSA-N L-methionine Chemical compound CSCC[C@H](N)C(O)=O FFEARJCKVFRZRR-BYPYZUCNSA-N 0.000 description 1
- 241000759772 Lambertia inermis Species 0.000 description 1
- 108091026898 Leader sequence (mRNA) Proteins 0.000 description 1
- CZCSUZMIRKFFFA-CIUDSAMLSA-N Leu-Ala-Asn Chemical compound [H]N[C@@H](CC(C)C)C(=O)N[C@@H](C)C(=O)N[C@@H](CC(N)=O)C(O)=O CZCSUZMIRKFFFA-CIUDSAMLSA-N 0.000 description 1
- DQPQTXMIRBUWKO-DCAQKATOSA-N Leu-Ala-Met Chemical compound C[C@@H](C(=O)N[C@@H](CCSC)C(=O)O)NC(=O)[C@H](CC(C)C)N DQPQTXMIRBUWKO-DCAQKATOSA-N 0.000 description 1
- XBBKIIGCUMBKCO-JXUBOQSCSA-N Leu-Ala-Thr Chemical compound [H]N[C@@H](CC(C)C)C(=O)N[C@@H](C)C(=O)N[C@@H]([C@@H](C)O)C(O)=O XBBKIIGCUMBKCO-JXUBOQSCSA-N 0.000 description 1
- HASRFYOMVPJRPU-SRVKXCTJSA-N Leu-Arg-Glu Chemical compound CC(C)C[C@H](N)C(=O)N[C@@H](CCCN=C(N)N)C(=O)N[C@@H](CCC(O)=O)C(O)=O HASRFYOMVPJRPU-SRVKXCTJSA-N 0.000 description 1
- DUBAVOVZNZKEQQ-AVGNSLFASA-N Leu-Arg-Val Chemical compound CC(C)C[C@H](N)C(=O)N[C@H](C(=O)N[C@@H](C(C)C)C(O)=O)CCCN=C(N)N DUBAVOVZNZKEQQ-AVGNSLFASA-N 0.000 description 1
- WUFYAPWIHCUMLL-CIUDSAMLSA-N Leu-Asn-Ala Chemical compound [H]N[C@@H](CC(C)C)C(=O)N[C@@H](CC(N)=O)C(=O)N[C@@H](C)C(O)=O WUFYAPWIHCUMLL-CIUDSAMLSA-N 0.000 description 1
- IGUOAYLTQJLPPD-DCAQKATOSA-N Leu-Asn-Arg Chemical compound CC(C)C[C@H](N)C(=O)N[C@@H](CC(N)=O)C(=O)N[C@H](C(O)=O)CCCN=C(N)N IGUOAYLTQJLPPD-DCAQKATOSA-N 0.000 description 1
- OIARJGNVARWKFP-YUMQZZPRSA-N Leu-Asn-Gly Chemical compound [H]N[C@@H](CC(C)C)C(=O)N[C@@H](CC(N)=O)C(=O)NCC(O)=O OIARJGNVARWKFP-YUMQZZPRSA-N 0.000 description 1
- MDVZJYGNAGLPGJ-KKUMJFAQSA-N Leu-Asn-Phe Chemical compound CC(C)C[C@H](N)C(=O)N[C@@H](CC(N)=O)C(=O)N[C@H](C(O)=O)CC1=CC=CC=C1 MDVZJYGNAGLPGJ-KKUMJFAQSA-N 0.000 description 1
- BPANDPNDMJHFEV-CIUDSAMLSA-N Leu-Asp-Ala Chemical compound [H]N[C@@H](CC(C)C)C(=O)N[C@@H](CC(O)=O)C(=O)N[C@@H](C)C(O)=O BPANDPNDMJHFEV-CIUDSAMLSA-N 0.000 description 1
- YKNBJXOJTURHCU-DCAQKATOSA-N Leu-Asp-Arg Chemical compound CC(C)C[C@H](N)C(=O)N[C@@H](CC(O)=O)C(=O)N[C@H](C(O)=O)CCCN=C(N)N YKNBJXOJTURHCU-DCAQKATOSA-N 0.000 description 1
- PJYSOYLLTJKZHC-GUBZILKMSA-N Leu-Asp-Gln Chemical compound CC(C)C[C@H](N)C(=O)N[C@@H](CC(O)=O)C(=O)N[C@H](C(O)=O)CCC(N)=O PJYSOYLLTJKZHC-GUBZILKMSA-N 0.000 description 1
- KTFHTMHHKXUYPW-ZPFDUUQYSA-N Leu-Asp-Ile Chemical compound [H]N[C@@H](CC(C)C)C(=O)N[C@@H](CC(O)=O)C(=O)N[C@@H]([C@@H](C)CC)C(O)=O KTFHTMHHKXUYPW-ZPFDUUQYSA-N 0.000 description 1
- MMEDVBWCMGRKKC-GARJFASQSA-N Leu-Asp-Pro Chemical compound CC(C)C[C@@H](C(=O)N[C@@H](CC(=O)O)C(=O)N1CCC[C@@H]1C(=O)O)N MMEDVBWCMGRKKC-GARJFASQSA-N 0.000 description 1
- VPKIQULSKFVCSM-SRVKXCTJSA-N Leu-Gln-Arg Chemical compound [H]N[C@@H](CC(C)C)C(=O)N[C@@H](CCC(N)=O)C(=O)N[C@@H](CCCNC(N)=N)C(O)=O VPKIQULSKFVCSM-SRVKXCTJSA-N 0.000 description 1
- LOLUPZNNADDTAA-AVGNSLFASA-N Leu-Gln-Leu Chemical compound CC(C)C[C@H](N)C(=O)N[C@@H](CCC(N)=O)C(=O)N[C@@H](CC(C)C)C(O)=O LOLUPZNNADDTAA-AVGNSLFASA-N 0.000 description 1
- FQZPTCNSNPWHLJ-AVGNSLFASA-N Leu-Gln-Lys Chemical compound CC(C)C[C@H](N)C(=O)N[C@@H](CCC(N)=O)C(=O)N[C@@H](CCCCN)C(O)=O FQZPTCNSNPWHLJ-AVGNSLFASA-N 0.000 description 1
- KUEVMUXNILMJTK-JYJNAYRXSA-N Leu-Gln-Tyr Chemical compound CC(C)C[C@H](N)C(=O)N[C@@H](CCC(N)=O)C(=O)N[C@H](C(O)=O)CC1=CC=C(O)C=C1 KUEVMUXNILMJTK-JYJNAYRXSA-N 0.000 description 1
- RVVBWTWPNFDYBE-SRVKXCTJSA-N Leu-Glu-Arg Chemical compound [H]N[C@@H](CC(C)C)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CCCNC(N)=N)C(O)=O RVVBWTWPNFDYBE-SRVKXCTJSA-N 0.000 description 1
- NEEOBPIXKWSBRF-IUCAKERBSA-N Leu-Glu-Gly Chemical compound [H]N[C@@H](CC(C)C)C(=O)N[C@@H](CCC(O)=O)C(=O)NCC(O)=O NEEOBPIXKWSBRF-IUCAKERBSA-N 0.000 description 1
- OGUUKPXUTHOIAV-SDDRHHMPSA-N Leu-Glu-Pro Chemical compound CC(C)C[C@@H](C(=O)N[C@@H](CCC(=O)O)C(=O)N1CCC[C@@H]1C(=O)O)N OGUUKPXUTHOIAV-SDDRHHMPSA-N 0.000 description 1
- HYIFFZAQXPUEAU-QWRGUYRKSA-N Leu-Gly-Leu Chemical compound CC(C)C[C@H](N)C(=O)NCC(=O)N[C@H](C(O)=O)CC(C)C HYIFFZAQXPUEAU-QWRGUYRKSA-N 0.000 description 1
- VZBIUJURDLFFOE-IHRRRGAJSA-N Leu-His-Arg Chemical compound [H]N[C@@H](CC(C)C)C(=O)N[C@@H](CC1=CNC=N1)C(=O)N[C@@H](CCCNC(N)=N)C(O)=O VZBIUJURDLFFOE-IHRRRGAJSA-N 0.000 description 1
- LIINDKYIGYTDLG-PPCPHDFISA-N Leu-Ile-Thr Chemical compound [H]N[C@@H](CC(C)C)C(=O)N[C@@H]([C@@H](C)CC)C(=O)N[C@@H]([C@@H](C)O)C(O)=O LIINDKYIGYTDLG-PPCPHDFISA-N 0.000 description 1
- YOKVEHGYYQEQOP-QWRGUYRKSA-N Leu-Leu-Gly Chemical compound CC(C)C[C@H](N)C(=O)N[C@@H](CC(C)C)C(=O)NCC(O)=O YOKVEHGYYQEQOP-QWRGUYRKSA-N 0.000 description 1
- XVZCXCTYGHPNEM-UHFFFAOYSA-N Leu-Leu-Pro Natural products CC(C)CC(N)C(=O)NC(CC(C)C)C(=O)N1CCCC1C(O)=O XVZCXCTYGHPNEM-UHFFFAOYSA-N 0.000 description 1
- WXUOJXIGOPMDJM-SRVKXCTJSA-N Leu-Lys-Asn Chemical compound CC(C)C[C@H](N)C(=O)N[C@@H](CCCCN)C(=O)N[C@@H](CC(N)=O)C(O)=O WXUOJXIGOPMDJM-SRVKXCTJSA-N 0.000 description 1
- HVHRPWQEQHIQJF-AVGNSLFASA-N Leu-Lys-Glu Chemical compound [H]N[C@@H](CC(C)C)C(=O)N[C@@H](CCCCN)C(=O)N[C@@H](CCC(O)=O)C(O)=O HVHRPWQEQHIQJF-AVGNSLFASA-N 0.000 description 1
- LZHJZLHSRGWBBE-IHRRRGAJSA-N Leu-Lys-Val Chemical compound [H]N[C@@H](CC(C)C)C(=O)N[C@@H](CCCCN)C(=O)N[C@@H](C(C)C)C(O)=O LZHJZLHSRGWBBE-IHRRRGAJSA-N 0.000 description 1
- INCJJHQRZGQLFC-KBPBESRZSA-N Leu-Phe-Gly Chemical compound [H]N[C@@H](CC(C)C)C(=O)N[C@@H](CC1=CC=CC=C1)C(=O)NCC(O)=O INCJJHQRZGQLFC-KBPBESRZSA-N 0.000 description 1
- RRVCZCNFXIFGRA-DCAQKATOSA-N Leu-Pro-Asn Chemical compound [H]N[C@@H](CC(C)C)C(=O)N1CCC[C@H]1C(=O)N[C@@H](CC(N)=O)C(O)=O RRVCZCNFXIFGRA-DCAQKATOSA-N 0.000 description 1
- UCBPDSYUVAAHCD-UWVGGRQHSA-N Leu-Pro-Gly Chemical compound CC(C)C[C@H](N)C(=O)N1CCC[C@H]1C(=O)NCC(O)=O UCBPDSYUVAAHCD-UWVGGRQHSA-N 0.000 description 1
- MUCIDQMDOYQYBR-IHRRRGAJSA-N Leu-Pro-His Chemical compound CC(C)C[C@@H](C(=O)N1CCC[C@H]1C(=O)N[C@@H](CC2=CN=CN2)C(=O)O)N MUCIDQMDOYQYBR-IHRRRGAJSA-N 0.000 description 1
- IDGZVZJLYFTXSL-DCAQKATOSA-N Leu-Ser-Arg Chemical compound CC(C)C[C@H](N)C(=O)N[C@@H](CO)C(=O)N[C@H](C(O)=O)CCCN=C(N)N IDGZVZJLYFTXSL-DCAQKATOSA-N 0.000 description 1
- JIHDFWWRYHSAQB-GUBZILKMSA-N Leu-Ser-Glu Chemical compound CC(C)C[C@H](N)C(=O)N[C@@H](CO)C(=O)N[C@H](C(O)=O)CCC(O)=O JIHDFWWRYHSAQB-GUBZILKMSA-N 0.000 description 1
- RGUXWMDNCPMQFB-YUMQZZPRSA-N Leu-Ser-Gly Chemical compound CC(C)C[C@H](N)C(=O)N[C@@H](CO)C(=O)NCC(O)=O RGUXWMDNCPMQFB-YUMQZZPRSA-N 0.000 description 1
- XOWMDXHFSBCAKQ-SRVKXCTJSA-N Leu-Ser-Leu Chemical compound CC(C)C[C@H](N)C(=O)N[C@@H](CO)C(=O)N[C@H](C(O)=O)CC(C)C XOWMDXHFSBCAKQ-SRVKXCTJSA-N 0.000 description 1
- AMSSKPUHBUQBOQ-SRVKXCTJSA-N Leu-Ser-Lys Chemical compound CC(C)C[C@@H](C(=O)N[C@@H](CO)C(=O)N[C@@H](CCCCN)C(=O)O)N AMSSKPUHBUQBOQ-SRVKXCTJSA-N 0.000 description 1
- SBANPBVRHYIMRR-UHFFFAOYSA-N Leu-Ser-Pro Natural products CC(C)CC(N)C(=O)NC(CO)C(=O)N1CCCC1C(O)=O SBANPBVRHYIMRR-UHFFFAOYSA-N 0.000 description 1
- BRTVHXHCUSXYRI-CIUDSAMLSA-N Leu-Ser-Ser Chemical compound CC(C)C[C@H](N)C(=O)N[C@@H](CO)C(=O)N[C@@H](CO)C(O)=O BRTVHXHCUSXYRI-CIUDSAMLSA-N 0.000 description 1
- SVBJIZVVYJYGLA-DCAQKATOSA-N Leu-Ser-Val Chemical compound [H]N[C@@H](CC(C)C)C(=O)N[C@@H](CO)C(=O)N[C@@H](C(C)C)C(O)=O SVBJIZVVYJYGLA-DCAQKATOSA-N 0.000 description 1
- ZDJQVSIPFLMNOX-RHYQMDGZSA-N Leu-Thr-Arg Chemical compound CC(C)C[C@H](N)C(=O)N[C@@H]([C@@H](C)O)C(=O)N[C@H](C(O)=O)CCCN=C(N)N ZDJQVSIPFLMNOX-RHYQMDGZSA-N 0.000 description 1
- LJBVRCDPWOJOEK-PPCPHDFISA-N Leu-Thr-Ile Chemical compound [H]N[C@@H](CC(C)C)C(=O)N[C@@H]([C@@H](C)O)C(=O)N[C@@H]([C@@H](C)CC)C(O)=O LJBVRCDPWOJOEK-PPCPHDFISA-N 0.000 description 1
- DAYQSYGBCUKVKT-VOAKCMCISA-N Leu-Thr-Lys Chemical compound CC(C)C[C@H](N)C(=O)N[C@@H]([C@@H](C)O)C(=O)N[C@@H](CCCCN)C(O)=O DAYQSYGBCUKVKT-VOAKCMCISA-N 0.000 description 1
- AIQWYVFNBNNOLU-RHYQMDGZSA-N Leu-Thr-Val Chemical compound [H]N[C@@H](CC(C)C)C(=O)N[C@@H]([C@@H](C)O)C(=O)N[C@@H](C(C)C)C(O)=O AIQWYVFNBNNOLU-RHYQMDGZSA-N 0.000 description 1
- WUHBLPVELFTPQK-KKUMJFAQSA-N Leu-Tyr-Asn Chemical compound [H]N[C@@H](CC(C)C)C(=O)N[C@@H](CC1=CC=C(O)C=C1)C(=O)N[C@@H](CC(N)=O)C(O)=O WUHBLPVELFTPQK-KKUMJFAQSA-N 0.000 description 1
- VJGQRELPQWNURN-JYJNAYRXSA-N Leu-Tyr-Glu Chemical compound [H]N[C@@H](CC(C)C)C(=O)N[C@@H](CC1=CC=C(O)C=C1)C(=O)N[C@@H](CCC(O)=O)C(O)=O VJGQRELPQWNURN-JYJNAYRXSA-N 0.000 description 1
- AAKRWBIIGKPOKQ-ONGXEEELSA-N Leu-Val-Gly Chemical compound CC(C)C[C@H](N)C(=O)N[C@@H](C(C)C)C(=O)NCC(O)=O AAKRWBIIGKPOKQ-ONGXEEELSA-N 0.000 description 1
- WSXTWLJHTLRFLW-SRVKXCTJSA-N Lys-Ala-Lys Chemical compound NCCCC[C@H](N)C(=O)N[C@@H](C)C(=O)N[C@@H](CCCCN)C(O)=O WSXTWLJHTLRFLW-SRVKXCTJSA-N 0.000 description 1
- IXHKPDJKKCUKHS-GARJFASQSA-N Lys-Ala-Pro Chemical compound C[C@@H](C(=O)N1CCC[C@@H]1C(=O)O)NC(=O)[C@H](CCCCN)N IXHKPDJKKCUKHS-GARJFASQSA-N 0.000 description 1
- WXJKFRMKJORORD-DCAQKATOSA-N Lys-Arg-Ala Chemical compound NC(=N)NCCC[C@@H](C(=O)N[C@@H](C)C(O)=O)NC(=O)[C@@H](N)CCCCN WXJKFRMKJORORD-DCAQKATOSA-N 0.000 description 1
- CLBGMWIYPYAZPR-AVGNSLFASA-N Lys-Arg-Arg Chemical compound NCCCC[C@H](N)C(=O)N[C@@H](CCCN=C(N)N)C(=O)N[C@@H](CCCN=C(N)N)C(O)=O CLBGMWIYPYAZPR-AVGNSLFASA-N 0.000 description 1
- ZTPWXNOOKAXPPE-DCAQKATOSA-N Lys-Arg-Cys Chemical compound C(CCN)C[C@@H](C(=O)N[C@@H](CCCN=C(N)N)C(=O)N[C@@H](CS)C(=O)O)N ZTPWXNOOKAXPPE-DCAQKATOSA-N 0.000 description 1
- WALVCOOOKULCQM-ULQDDVLXSA-N Lys-Arg-Phe Chemical compound [H]N[C@@H](CCCCN)C(=O)N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CC1=CC=CC=C1)C(O)=O WALVCOOOKULCQM-ULQDDVLXSA-N 0.000 description 1
- DNEJSAIMVANNPA-DCAQKATOSA-N Lys-Asn-Arg Chemical compound [H]N[C@@H](CCCCN)C(=O)N[C@@H](CC(N)=O)C(=O)N[C@@H](CCCNC(N)=N)C(O)=O DNEJSAIMVANNPA-DCAQKATOSA-N 0.000 description 1
- YKIRNDPUWONXQN-GUBZILKMSA-N Lys-Asn-Gln Chemical compound C(CCN)C[C@@H](C(=O)N[C@@H](CC(=O)N)C(=O)N[C@@H](CCC(=O)N)C(=O)O)N YKIRNDPUWONXQN-GUBZILKMSA-N 0.000 description 1
- IWWMPCPLFXFBAF-SRVKXCTJSA-N Lys-Asp-Leu Chemical compound [H]N[C@@H](CCCCN)C(=O)N[C@@H](CC(O)=O)C(=O)N[C@@H](CC(C)C)C(O)=O IWWMPCPLFXFBAF-SRVKXCTJSA-N 0.000 description 1
- QIJVAFLRMVBHMU-KKUMJFAQSA-N Lys-Asp-Phe Chemical compound [H]N[C@@H](CCCCN)C(=O)N[C@@H](CC(O)=O)C(=O)N[C@@H](CC1=CC=CC=C1)C(O)=O QIJVAFLRMVBHMU-KKUMJFAQSA-N 0.000 description 1
- DFXQCCBKGUNYGG-GUBZILKMSA-N Lys-Gln-Ala Chemical compound OC(=O)[C@H](C)NC(=O)[C@H](CCC(N)=O)NC(=O)[C@@H](N)CCCCN DFXQCCBKGUNYGG-GUBZILKMSA-N 0.000 description 1
- QQUJSUFWEDZQQY-AVGNSLFASA-N Lys-Gln-Lys Chemical compound NCCCC[C@H](N)C(=O)N[C@@H](CCC(N)=O)C(=O)N[C@H](C(O)=O)CCCCN QQUJSUFWEDZQQY-AVGNSLFASA-N 0.000 description 1
- NNCDAORZCMPZPX-GUBZILKMSA-N Lys-Gln-Ser Chemical compound C(CCN)C[C@@H](C(=O)N[C@@H](CCC(=O)N)C(=O)N[C@@H](CO)C(=O)O)N NNCDAORZCMPZPX-GUBZILKMSA-N 0.000 description 1
- VEGLGAOVLFODGC-GUBZILKMSA-N Lys-Glu-Ser Chemical compound [H]N[C@@H](CCCCN)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CO)C(O)=O VEGLGAOVLFODGC-GUBZILKMSA-N 0.000 description 1
- DTUZCYRNEJDKSR-NHCYSSNCSA-N Lys-Gly-Ile Chemical compound CC[C@H](C)[C@@H](C(O)=O)NC(=O)CNC(=O)[C@@H](N)CCCCN DTUZCYRNEJDKSR-NHCYSSNCSA-N 0.000 description 1
- PBLLTSKBTAHDNA-KBPBESRZSA-N Lys-Gly-Phe Chemical compound [H]N[C@@H](CCCCN)C(=O)NCC(=O)N[C@@H](CC1=CC=CC=C1)C(O)=O PBLLTSKBTAHDNA-KBPBESRZSA-N 0.000 description 1
- FHIAJWBDZVHLAH-YUMQZZPRSA-N Lys-Gly-Ser Chemical compound NCCCC[C@H](N)C(=O)NCC(=O)N[C@@H](CO)C(O)=O FHIAJWBDZVHLAH-YUMQZZPRSA-N 0.000 description 1
- KZJQUYFDSCFSCO-DLOVCJGASA-N Lys-His-Ala Chemical compound C[C@@H](C(=O)O)NC(=O)[C@H](CC1=CN=CN1)NC(=O)[C@H](CCCCN)N KZJQUYFDSCFSCO-DLOVCJGASA-N 0.000 description 1
- SLQJJFAVWSZLBL-BJDJZHNGSA-N Lys-Ile-Ala Chemical compound OC(=O)[C@H](C)NC(=O)[C@H]([C@@H](C)CC)NC(=O)[C@@H](N)CCCCN SLQJJFAVWSZLBL-BJDJZHNGSA-N 0.000 description 1
- MXMDJEJWERYPMO-XUXIUFHCSA-N Lys-Ile-Arg Chemical compound [H]N[C@@H](CCCCN)C(=O)N[C@@H]([C@@H](C)CC)C(=O)N[C@@H](CCCNC(N)=N)C(O)=O MXMDJEJWERYPMO-XUXIUFHCSA-N 0.000 description 1
- XREQQOATSMMAJP-MGHWNKPDSA-N Lys-Ile-Tyr Chemical compound [H]N[C@@H](CCCCN)C(=O)N[C@@H]([C@@H](C)CC)C(=O)N[C@@H](CC1=CC=C(O)C=C1)C(O)=O XREQQOATSMMAJP-MGHWNKPDSA-N 0.000 description 1
- NJNRBRKHOWSGMN-SRVKXCTJSA-N Lys-Leu-Asn Chemical compound [H]N[C@@H](CCCCN)C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](CC(N)=O)C(O)=O NJNRBRKHOWSGMN-SRVKXCTJSA-N 0.000 description 1
- SKRGVGLIRUGANF-AVGNSLFASA-N Lys-Leu-Glu Chemical compound [H]N[C@@H](CCCCN)C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](CCC(O)=O)C(O)=O SKRGVGLIRUGANF-AVGNSLFASA-N 0.000 description 1
- RIJCHEVHFWMDKD-SRVKXCTJSA-N Lys-Lys-Asn Chemical compound NCCCC[C@H](N)C(=O)N[C@@H](CCCCN)C(=O)N[C@@H](CC(N)=O)C(O)=O RIJCHEVHFWMDKD-SRVKXCTJSA-N 0.000 description 1
- YUAXTFMFMOIMAM-QWRGUYRKSA-N Lys-Lys-Gly Chemical compound [H]N[C@@H](CCCCN)C(=O)N[C@@H](CCCCN)C(=O)NCC(O)=O YUAXTFMFMOIMAM-QWRGUYRKSA-N 0.000 description 1
- PIXVFCBYEGPZPA-JYJNAYRXSA-N Lys-Phe-Gln Chemical compound C1=CC=C(C=C1)C[C@@H](C(=O)N[C@@H](CCC(=O)N)C(=O)O)NC(=O)[C@H](CCCCN)N PIXVFCBYEGPZPA-JYJNAYRXSA-N 0.000 description 1
- ZJSZPXISKMDJKQ-JYJNAYRXSA-N Lys-Phe-Glu Chemical compound NCCCC[C@H](N)C(=O)N[C@H](C(=O)N[C@@H](CCC(O)=O)C(O)=O)CC1=CC=CC=C1 ZJSZPXISKMDJKQ-JYJNAYRXSA-N 0.000 description 1
- OBZHNHBAAVEWKI-DCAQKATOSA-N Lys-Pro-Asn Chemical compound NCCCC[C@H](N)C(=O)N1CCC[C@H]1C(=O)N[C@@H](CC(N)=O)C(O)=O OBZHNHBAAVEWKI-DCAQKATOSA-N 0.000 description 1
- WGILOYIKJVQUPT-DCAQKATOSA-N Lys-Pro-Asp Chemical compound [H]N[C@@H](CCCCN)C(=O)N1CCC[C@H]1C(=O)N[C@@H](CC(O)=O)C(O)=O WGILOYIKJVQUPT-DCAQKATOSA-N 0.000 description 1
- LECIJRIRMVOFMH-ULQDDVLXSA-N Lys-Pro-Phe Chemical compound NCCCC[C@H](N)C(=O)N1CCC[C@H]1C(=O)N[C@H](C(O)=O)CC1=CC=CC=C1 LECIJRIRMVOFMH-ULQDDVLXSA-N 0.000 description 1
- HKXSZKJMDBHOTG-CIUDSAMLSA-N Lys-Ser-Ala Chemical compound OC(=O)[C@H](C)NC(=O)[C@H](CO)NC(=O)[C@@H](N)CCCCN HKXSZKJMDBHOTG-CIUDSAMLSA-N 0.000 description 1
- SBQDRNOLGSYHQA-YUMQZZPRSA-N Lys-Ser-Gly Chemical compound [H]N[C@@H](CCCCN)C(=O)N[C@@H](CO)C(=O)NCC(O)=O SBQDRNOLGSYHQA-YUMQZZPRSA-N 0.000 description 1
- JOSAKOKSPXROGQ-BJDJZHNGSA-N Lys-Ser-Ile Chemical compound [H]N[C@@H](CCCCN)C(=O)N[C@@H](CO)C(=O)N[C@@H]([C@@H](C)CC)C(O)=O JOSAKOKSPXROGQ-BJDJZHNGSA-N 0.000 description 1
- WZVSHTFTCYOFPL-GARJFASQSA-N Lys-Ser-Pro Chemical compound C1C[C@@H](N(C1)C(=O)[C@H](CO)NC(=O)[C@H](CCCCN)N)C(=O)O WZVSHTFTCYOFPL-GARJFASQSA-N 0.000 description 1
- DIBZLYZXTSVGLN-CIUDSAMLSA-N Lys-Ser-Ser Chemical compound [H]N[C@@H](CCCCN)C(=O)N[C@@H](CO)C(=O)N[C@@H](CO)C(O)=O DIBZLYZXTSVGLN-CIUDSAMLSA-N 0.000 description 1
- TVHCDSBMFQYPNA-RHYQMDGZSA-N Lys-Thr-Arg Chemical compound [H]N[C@@H](CCCCN)C(=O)N[C@@H]([C@@H](C)O)C(=O)N[C@@H](CCCNC(N)=N)C(O)=O TVHCDSBMFQYPNA-RHYQMDGZSA-N 0.000 description 1
- XATKLFSXFINPSB-JYJNAYRXSA-N Lys-Tyr-Gln Chemical compound [H]N[C@@H](CCCCN)C(=O)N[C@@H](CC1=CC=C(O)C=C1)C(=O)N[C@@H](CCC(N)=O)C(O)=O XATKLFSXFINPSB-JYJNAYRXSA-N 0.000 description 1
- RQILLQOQXLZTCK-KBPBESRZSA-N Lys-Tyr-Gly Chemical compound [H]N[C@@H](CCCCN)C(=O)N[C@@H](CC1=CC=C(O)C=C1)C(=O)NCC(O)=O RQILLQOQXLZTCK-KBPBESRZSA-N 0.000 description 1
- TXTZMVNJIRZABH-ULQDDVLXSA-N Lys-Val-Phe Chemical compound NCCCC[C@H](N)C(=O)N[C@@H](C(C)C)C(=O)N[C@H](C(O)=O)CC1=CC=CC=C1 TXTZMVNJIRZABH-ULQDDVLXSA-N 0.000 description 1
- 101000930511 Macadamia integrifolia Antimicrobial peptide 1 Proteins 0.000 description 1
- 241000124008 Mammalia Species 0.000 description 1
- 101000763602 Manilkara zapota Thaumatin-like protein 1 Proteins 0.000 description 1
- 101000763586 Manilkara zapota Thaumatin-like protein 1a Proteins 0.000 description 1
- HUKLXYYPZWPXCC-KZVJFYERSA-N Met-Ala-Thr Chemical compound CSCC[C@H](N)C(=O)N[C@@H](C)C(=O)N[C@@H]([C@@H](C)O)C(O)=O HUKLXYYPZWPXCC-KZVJFYERSA-N 0.000 description 1
- DSWOTZCVCBEPOU-IUCAKERBSA-N Met-Arg-Gly Chemical compound CSCC[C@H](N)C(=O)N[C@H](C(=O)NCC(O)=O)CCCNC(N)=N DSWOTZCVCBEPOU-IUCAKERBSA-N 0.000 description 1
- IYXDSYWCVVXSKB-CIUDSAMLSA-N Met-Asn-Glu Chemical compound [H]N[C@@H](CCSC)C(=O)N[C@@H](CC(N)=O)C(=O)N[C@@H](CCC(O)=O)C(O)=O IYXDSYWCVVXSKB-CIUDSAMLSA-N 0.000 description 1
- IHITVQKJXQQGLJ-LPEHRKFASA-N Met-Asn-Pro Chemical compound CSCC[C@@H](C(=O)N[C@@H](CC(=O)N)C(=O)N1CCC[C@@H]1C(=O)O)N IHITVQKJXQQGLJ-LPEHRKFASA-N 0.000 description 1
- OOSPRDCGTLQLBP-NHCYSSNCSA-N Met-Glu-Val Chemical compound [H]N[C@@H](CCSC)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](C(C)C)C(O)=O OOSPRDCGTLQLBP-NHCYSSNCSA-N 0.000 description 1
- XKJUFUPCHARJKX-UWVGGRQHSA-N Met-Gly-His Chemical compound CSCC[C@H](N)C(=O)NCC(=O)N[C@H](C(O)=O)CC1=CNC=N1 XKJUFUPCHARJKX-UWVGGRQHSA-N 0.000 description 1
- AWGBEIYZPAXXSX-RWMBFGLXSA-N Met-Leu-Pro Chemical compound CC(C)C[C@@H](C(=O)N1CCC[C@@H]1C(=O)O)NC(=O)[C@H](CCSC)N AWGBEIYZPAXXSX-RWMBFGLXSA-N 0.000 description 1
- OXIWIYOJVNOKOV-SRVKXCTJSA-N Met-Met-Arg Chemical compound CSCC[C@H](N)C(=O)N[C@@H](CCSC)C(=O)N[C@H](C(O)=O)CCCNC(N)=N OXIWIYOJVNOKOV-SRVKXCTJSA-N 0.000 description 1
- JOYFULUKJRJCSX-IUCAKERBSA-N Met-Met-Gly Chemical compound CSCC[C@H](N)C(=O)N[C@@H](CCSC)C(=O)NCC(O)=O JOYFULUKJRJCSX-IUCAKERBSA-N 0.000 description 1
- WUYLWZRHRLLEGB-AVGNSLFASA-N Met-Met-Leu Chemical compound CSCC[C@H](N)C(=O)N[C@@H](CCSC)C(=O)N[C@@H](CC(C)C)C(O)=O WUYLWZRHRLLEGB-AVGNSLFASA-N 0.000 description 1
- LNXGEYIEEUZGGH-JYJNAYRXSA-N Met-Phe-Arg Chemical compound NC(N)=NCCC[C@@H](C(O)=O)NC(=O)[C@@H](NC(=O)[C@@H](N)CCSC)CC1=CC=CC=C1 LNXGEYIEEUZGGH-JYJNAYRXSA-N 0.000 description 1
- VQILILSLEFDECU-GUBZILKMSA-N Met-Pro-Ala Chemical compound [H]N[C@@H](CCSC)C(=O)N1CCC[C@H]1C(=O)N[C@@H](C)C(O)=O VQILILSLEFDECU-GUBZILKMSA-N 0.000 description 1
- GMMLGMFBYCFCCX-KZVJFYERSA-N Met-Thr-Ala Chemical compound CSCC[C@H](N)C(=O)N[C@@H]([C@@H](C)O)C(=O)N[C@@H](C)C(O)=O GMMLGMFBYCFCCX-KZVJFYERSA-N 0.000 description 1
- LBSWWNKMVPAXOI-GUBZILKMSA-N Met-Val-Ser Chemical compound CSCC[C@H](N)C(=O)N[C@@H](C(C)C)C(=O)N[C@@H](CO)C(O)=O LBSWWNKMVPAXOI-GUBZILKMSA-N 0.000 description 1
- IIHMNTBFPMRJCN-RCWTZXSCSA-N Met-Val-Thr Chemical compound CSCC[C@H](N)C(=O)N[C@@H](C(C)C)C(=O)N[C@@H]([C@@H](C)O)C(O)=O IIHMNTBFPMRJCN-RCWTZXSCSA-N 0.000 description 1
- -1 MiAMP2a and MiAMP2c) Chemical compound 0.000 description 1
- 241000713869 Moloney murine leukemia virus Species 0.000 description 1
- 101000966653 Musa acuminata Glucan endo-1,3-beta-glucosidase Proteins 0.000 description 1
- WYBVBIHNJWOLCJ-UHFFFAOYSA-N N-L-arginyl-L-leucine Natural products CC(C)CC(C(O)=O)NC(=O)C(N)CCCN=C(N)N WYBVBIHNJWOLCJ-UHFFFAOYSA-N 0.000 description 1
- AUEJLPRZGVVDNU-UHFFFAOYSA-N N-L-tyrosyl-L-leucine Natural products CC(C)CC(C(O)=O)NC(=O)C(N)CC1=CC=C(O)C=C1 AUEJLPRZGVVDNU-UHFFFAOYSA-N 0.000 description 1
- 229910004619 Na2MoO4 Inorganic materials 0.000 description 1
- 101100068676 Neurospora crassa (strain ATCC 24698 / 74-OR23-1A / CBS 708.71 / DSM 1257 / FGSC 987) gln-1 gene Proteins 0.000 description 1
- 241000208125 Nicotiana Species 0.000 description 1
- 241000224778 Nitraria billardierei Species 0.000 description 1
- 239000000020 Nitrocellulose Substances 0.000 description 1
- 241001547399 Petrophile canescens Species 0.000 description 1
- 108010002747 Pfu DNA polymerase Proteins 0.000 description 1
- DFEVBOYEUQJGER-JURCDPSOSA-N Phe-Ala-Ile Chemical compound N[C@@H](CC1=CC=CC=C1)C(=O)N[C@@H](C)C(=O)N[C@@H]([C@@H](C)CC)C(=O)O DFEVBOYEUQJGER-JURCDPSOSA-N 0.000 description 1
- DPUOLKQSMYLRDR-UBHSHLNASA-N Phe-Arg-Ala Chemical compound NC(N)=NCCC[C@@H](C(=O)N[C@@H](C)C(O)=O)NC(=O)[C@@H](N)CC1=CC=CC=C1 DPUOLKQSMYLRDR-UBHSHLNASA-N 0.000 description 1
- LZDIENNKWVXJMX-JYJNAYRXSA-N Phe-Arg-Arg Chemical compound NC(N)=NCCC[C@@H](C(O)=O)NC(=O)[C@H](CCCN=C(N)N)NC(=O)[C@@H](N)CC1=CC=CC=C1 LZDIENNKWVXJMX-JYJNAYRXSA-N 0.000 description 1
- XWBJLKDCHJVKAK-KKUMJFAQSA-N Phe-Arg-Gln Chemical compound C1=CC=C(C=C1)C[C@@H](C(=O)N[C@@H](CCCN=C(N)N)C(=O)N[C@@H](CCC(=O)N)C(=O)O)N XWBJLKDCHJVKAK-KKUMJFAQSA-N 0.000 description 1
- MPGJIHFJCXTVEX-KKUMJFAQSA-N Phe-Arg-Glu Chemical compound [H]N[C@@H](CC1=CC=CC=C1)C(=O)N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CCC(O)=O)C(O)=O MPGJIHFJCXTVEX-KKUMJFAQSA-N 0.000 description 1
- CGOMLCQJEMWMCE-STQMWFEESA-N Phe-Arg-Gly Chemical compound NC(N)=NCCC[C@@H](C(=O)NCC(O)=O)NC(=O)[C@@H](N)CC1=CC=CC=C1 CGOMLCQJEMWMCE-STQMWFEESA-N 0.000 description 1
- AGYXCMYVTBYGCT-ULQDDVLXSA-N Phe-Arg-Leu Chemical compound [H]N[C@@H](CC1=CC=CC=C1)C(=O)N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CC(C)C)C(O)=O AGYXCMYVTBYGCT-ULQDDVLXSA-N 0.000 description 1
- MECSIDWUTYRHRJ-KKUMJFAQSA-N Phe-Asn-Leu Chemical compound [H]N[C@@H](CC1=CC=CC=C1)C(=O)N[C@@H](CC(N)=O)C(=O)N[C@@H](CC(C)C)C(O)=O MECSIDWUTYRHRJ-KKUMJFAQSA-N 0.000 description 1
- HTKNPQZCMLBOTQ-XVSYOHENSA-N Phe-Asn-Thr Chemical compound C[C@H]([C@@H](C(=O)O)NC(=O)[C@H](CC(=O)N)NC(=O)[C@H](CC1=CC=CC=C1)N)O HTKNPQZCMLBOTQ-XVSYOHENSA-N 0.000 description 1
- UEEVBGHEGJMDDV-AVGNSLFASA-N Phe-Asp-Gln Chemical compound NC(=O)CC[C@@H](C(O)=O)NC(=O)[C@H](CC(O)=O)NC(=O)[C@@H](N)CC1=CC=CC=C1 UEEVBGHEGJMDDV-AVGNSLFASA-N 0.000 description 1
- DDYIRGBOZVKRFR-AVGNSLFASA-N Phe-Asp-Glu Chemical compound C1=CC=C(C=C1)C[C@@H](C(=O)N[C@@H](CC(=O)O)C(=O)N[C@@H](CCC(=O)O)C(=O)O)N DDYIRGBOZVKRFR-AVGNSLFASA-N 0.000 description 1
- OPEVYHFJXLCCRT-AVGNSLFASA-N Phe-Gln-Ser Chemical compound [H]N[C@@H](CC1=CC=CC=C1)C(=O)N[C@@H](CCC(N)=O)C(=O)N[C@@H](CO)C(O)=O OPEVYHFJXLCCRT-AVGNSLFASA-N 0.000 description 1
- MGBRZXXGQBAULP-DRZSPHRISA-N Phe-Glu-Ala Chemical compound OC(=O)[C@H](C)NC(=O)[C@H](CCC(O)=O)NC(=O)[C@@H](N)CC1=CC=CC=C1 MGBRZXXGQBAULP-DRZSPHRISA-N 0.000 description 1
- MPFGIYLYWUCSJG-AVGNSLFASA-N Phe-Glu-Asp Chemical compound OC(=O)C[C@@H](C(O)=O)NC(=O)[C@H](CCC(O)=O)NC(=O)[C@@H](N)CC1=CC=CC=C1 MPFGIYLYWUCSJG-AVGNSLFASA-N 0.000 description 1
- CDQCFGOQNYOICK-IHRRRGAJSA-N Phe-Glu-Gln Chemical compound NC(=O)CC[C@@H](C(O)=O)NC(=O)[C@H](CCC(O)=O)NC(=O)[C@@H](N)CC1=CC=CC=C1 CDQCFGOQNYOICK-IHRRRGAJSA-N 0.000 description 1
- KYYMILWEGJYPQZ-IHRRRGAJSA-N Phe-Glu-Glu Chemical compound OC(=O)CC[C@@H](C(O)=O)NC(=O)[C@H](CCC(O)=O)NC(=O)[C@@H](N)CC1=CC=CC=C1 KYYMILWEGJYPQZ-IHRRRGAJSA-N 0.000 description 1
- MGECUMGTSHYHEJ-QEWYBTABSA-N Phe-Glu-Ile Chemical compound CC[C@H](C)[C@@H](C(O)=O)NC(=O)[C@H](CCC(O)=O)NC(=O)[C@@H](N)CC1=CC=CC=C1 MGECUMGTSHYHEJ-QEWYBTABSA-N 0.000 description 1
- BFYHIHGIHGROAT-HTUGSXCWSA-N Phe-Glu-Thr Chemical compound [H]N[C@@H](CC1=CC=CC=C1)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H]([C@@H](C)O)C(O)=O BFYHIHGIHGROAT-HTUGSXCWSA-N 0.000 description 1
- YYKZDTVQHTUKDW-RYUDHWBXSA-N Phe-Gly-Gln Chemical compound C1=CC=C(C=C1)C[C@@H](C(=O)NCC(=O)N[C@@H](CCC(=O)N)C(=O)O)N YYKZDTVQHTUKDW-RYUDHWBXSA-N 0.000 description 1
- NAXPHWZXEXNDIW-JTQLQIEISA-N Phe-Gly-Gly Chemical compound OC(=O)CNC(=O)CNC(=O)[C@@H](N)CC1=CC=CC=C1 NAXPHWZXEXNDIW-JTQLQIEISA-N 0.000 description 1
- HGNGAMWHGGANAU-WHOFXGATSA-N Phe-Gly-Ile Chemical compound CC[C@H](C)[C@@H](C(O)=O)NC(=O)CNC(=O)[C@@H](N)CC1=CC=CC=C1 HGNGAMWHGGANAU-WHOFXGATSA-N 0.000 description 1
- NPLGQVKZFGJWAI-QWHCGFSZSA-N Phe-Gly-Pro Chemical compound C1C[C@@H](N(C1)C(=O)CNC(=O)[C@H](CC2=CC=CC=C2)N)C(=O)O NPLGQVKZFGJWAI-QWHCGFSZSA-N 0.000 description 1
- KBVJZCVLQWCJQN-KKUMJFAQSA-N Phe-Leu-Asn Chemical compound [H]N[C@@H](CC1=CC=CC=C1)C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](CC(N)=O)C(O)=O KBVJZCVLQWCJQN-KKUMJFAQSA-N 0.000 description 1
- MSHZERMPZKCODG-ACRUOGEOSA-N Phe-Leu-Phe Chemical compound C([C@H](N)C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](CC=1C=CC=CC=1)C(O)=O)C1=CC=CC=C1 MSHZERMPZKCODG-ACRUOGEOSA-N 0.000 description 1
- ZUQACJLOHYRVPJ-DKIMLUQUSA-N Phe-Lys-Ile Chemical compound CC[C@H](C)[C@@H](C(O)=O)NC(=O)[C@H](CCCCN)NC(=O)[C@@H](N)CC1=CC=CC=C1 ZUQACJLOHYRVPJ-DKIMLUQUSA-N 0.000 description 1
- WEDZFLRYSIDIRX-IHRRRGAJSA-N Phe-Ser-Arg Chemical compound NC(=N)NCCC[C@@H](C(O)=O)NC(=O)[C@H](CO)NC(=O)[C@@H](N)CC1=CC=CC=C1 WEDZFLRYSIDIRX-IHRRRGAJSA-N 0.000 description 1
- UNBFGVQVQGXXCK-KKUMJFAQSA-N Phe-Ser-Leu Chemical compound [H]N[C@@H](CC1=CC=CC=C1)C(=O)N[C@@H](CO)C(=O)N[C@@H](CC(C)C)C(O)=O UNBFGVQVQGXXCK-KKUMJFAQSA-N 0.000 description 1
- RAGOJJCBGXARPO-XVSYOHENSA-N Phe-Thr-Asp Chemical compound OC(=O)C[C@@H](C(O)=O)NC(=O)[C@H]([C@H](O)C)NC(=O)[C@@H](N)CC1=CC=CC=C1 RAGOJJCBGXARPO-XVSYOHENSA-N 0.000 description 1
- CDHURCQGUDNBMA-UBHSHLNASA-N Phe-Val-Ala Chemical compound OC(=O)[C@H](C)NC(=O)[C@H](C(C)C)NC(=O)[C@@H](N)CC1=CC=CC=C1 CDHURCQGUDNBMA-UBHSHLNASA-N 0.000 description 1
- BQMFWUKNOCJDNV-HJWJTTGWSA-N Phe-Val-Ile Chemical compound [H]N[C@@H](CC1=CC=CC=C1)C(=O)N[C@@H](C(C)C)C(=O)N[C@@H]([C@@H](C)CC)C(O)=O BQMFWUKNOCJDNV-HJWJTTGWSA-N 0.000 description 1
- 241000233620 Phytophthora cryptogea Species 0.000 description 1
- 241000233629 Phytophthora parasitica Species 0.000 description 1
- 108010064851 Plant Proteins Proteins 0.000 description 1
- 108010021757 Polynucleotide 5'-Hydroxyl-Kinase Proteins 0.000 description 1
- 102000008422 Polynucleotide 5'-hydroxyl-kinase Human genes 0.000 description 1
- LNLNHXIQPGKRJQ-SRVKXCTJSA-N Pro-Arg-Arg Chemical compound NC(N)=NCCC[C@@H](C(O)=O)NC(=O)[C@H](CCCN=C(N)N)NC(=O)[C@@H]1CCCN1 LNLNHXIQPGKRJQ-SRVKXCTJSA-N 0.000 description 1
- QSKCKTUQPICLSO-AVGNSLFASA-N Pro-Arg-Lys Chemical compound C1C[C@H](NC1)C(=O)N[C@@H](CCCN=C(N)N)C(=O)N[C@@H](CCCCN)C(=O)O QSKCKTUQPICLSO-AVGNSLFASA-N 0.000 description 1
- ICTZKEXYDDZZFP-SRVKXCTJSA-N Pro-Arg-Pro Chemical compound N([C@@H](CCCN=C(N)N)C(=O)N1[C@@H](CCC1)C(O)=O)C(=O)[C@@H]1CCCN1 ICTZKEXYDDZZFP-SRVKXCTJSA-N 0.000 description 1
- UVKNEILZSJMKSR-FXQIFTODSA-N Pro-Asn-Ala Chemical compound OC(=O)[C@H](C)NC(=O)[C@H](CC(N)=O)NC(=O)[C@@H]1CCCN1 UVKNEILZSJMKSR-FXQIFTODSA-N 0.000 description 1
- CJZTUKSFZUSNCC-FXQIFTODSA-N Pro-Asp-Asn Chemical compound NC(=O)C[C@@H](C(O)=O)NC(=O)[C@H](CC(O)=O)NC(=O)[C@@H]1CCCN1 CJZTUKSFZUSNCC-FXQIFTODSA-N 0.000 description 1
- KPDRZQUWJKTMBP-DCAQKATOSA-N Pro-Asp-Leu Chemical compound CC(C)C[C@@H](C(=O)O)NC(=O)[C@H](CC(=O)O)NC(=O)[C@@H]1CCCN1 KPDRZQUWJKTMBP-DCAQKATOSA-N 0.000 description 1
- SNIPWBQKOPCJRG-CIUDSAMLSA-N Pro-Gln-Cys Chemical compound C1C[C@H](NC1)C(=O)N[C@@H](CCC(=O)N)C(=O)N[C@@H](CS)C(=O)O SNIPWBQKOPCJRG-CIUDSAMLSA-N 0.000 description 1
- HJSCRFZVGXAGNG-SRVKXCTJSA-N Pro-Gln-Leu Chemical compound CC(C)C[C@@H](C(O)=O)NC(=O)[C@H](CCC(N)=O)NC(=O)[C@@H]1CCCN1 HJSCRFZVGXAGNG-SRVKXCTJSA-N 0.000 description 1
- SKICPQLTOXGWGO-GARJFASQSA-N Pro-Gln-Pro Chemical compound C1C[C@H](NC1)C(=O)N[C@@H](CCC(=O)N)C(=O)N2CCC[C@@H]2C(=O)O SKICPQLTOXGWGO-GARJFASQSA-N 0.000 description 1
- FRKBNXCFJBPJOL-GUBZILKMSA-N Pro-Glu-Glu Chemical compound [H]N1CCC[C@H]1C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CCC(O)=O)C(O)=O FRKBNXCFJBPJOL-GUBZILKMSA-N 0.000 description 1
- VOZIBWWZSBIXQN-SRVKXCTJSA-N Pro-Glu-Lys Chemical compound NCCCC[C@H](NC(=O)[C@H](CCC(O)=O)NC(=O)[C@@H]1CCCN1)C(O)=O VOZIBWWZSBIXQN-SRVKXCTJSA-N 0.000 description 1
- LXVLKXPFIDDHJG-CIUDSAMLSA-N Pro-Glu-Ser Chemical compound [H]N1CCC[C@H]1C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CO)C(O)=O LXVLKXPFIDDHJG-CIUDSAMLSA-N 0.000 description 1
- VYWNORHENYEQDW-YUMQZZPRSA-N Pro-Gly-Glu Chemical compound OC(=O)CC[C@@H](C(O)=O)NC(=O)CNC(=O)[C@@H]1CCCN1 VYWNORHENYEQDW-YUMQZZPRSA-N 0.000 description 1
- UIMCLYYSUCIUJM-UWVGGRQHSA-N Pro-Gly-Lys Chemical compound NCCCC[C@@H](C(O)=O)NC(=O)CNC(=O)[C@@H]1CCCN1 UIMCLYYSUCIUJM-UWVGGRQHSA-N 0.000 description 1
- JUJGNDZIKKQMDJ-IHRRRGAJSA-N Pro-His-His Chemical compound [H]N1CCC[C@H]1C(=O)N[C@@H](CC1=CNC=N1)C(=O)N[C@@H](CC1=CNC=N1)C(O)=O JUJGNDZIKKQMDJ-IHRRRGAJSA-N 0.000 description 1
- FKVNLUZHSFCNGY-RVMXOQNASA-N Pro-Ile-Pro Chemical compound CC[C@H](C)[C@@H](C(=O)N1CCC[C@@H]1C(=O)O)NC(=O)[C@@H]2CCCN2 FKVNLUZHSFCNGY-RVMXOQNASA-N 0.000 description 1
- RYJRPPUATSKNAY-STECZYCISA-N Pro-Ile-Tyr Chemical compound CC[C@H](C)[C@@H](C(=O)N[C@@H](CC1=CC=C(C=C1)O)C(=O)O)NC(=O)[C@@H]2CCCN2 RYJRPPUATSKNAY-STECZYCISA-N 0.000 description 1
- HFNPOYOKIPGAEI-SRVKXCTJSA-N Pro-Leu-Glu Chemical compound OC(=O)CC[C@@H](C(O)=O)NC(=O)[C@H](CC(C)C)NC(=O)[C@@H]1CCCN1 HFNPOYOKIPGAEI-SRVKXCTJSA-N 0.000 description 1
- HATVCTYBNCNMAA-AVGNSLFASA-N Pro-Leu-Met Chemical compound [H]N1CCC[C@H]1C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](CCSC)C(O)=O HATVCTYBNCNMAA-AVGNSLFASA-N 0.000 description 1
- FKYKZHOKDOPHSA-DCAQKATOSA-N Pro-Leu-Ser Chemical compound [H]N1CCC[C@H]1C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](CO)C(O)=O FKYKZHOKDOPHSA-DCAQKATOSA-N 0.000 description 1
- SUENWIFTSTWUKD-AVGNSLFASA-N Pro-Leu-Val Chemical compound [H]N1CCC[C@H]1C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](C(C)C)C(O)=O SUENWIFTSTWUKD-AVGNSLFASA-N 0.000 description 1
- CDGABSWLRMECHC-IHRRRGAJSA-N Pro-Lys-His Chemical compound C1C[C@H](NC1)C(=O)N[C@@H](CCCCN)C(=O)N[C@@H](CC2=CN=CN2)C(=O)O CDGABSWLRMECHC-IHRRRGAJSA-N 0.000 description 1
- DWGFLKQSGRUQTI-IHRRRGAJSA-N Pro-Lys-Lys Chemical compound NCCCC[C@@H](C(O)=O)NC(=O)[C@H](CCCCN)NC(=O)[C@@H]1CCCN1 DWGFLKQSGRUQTI-IHRRRGAJSA-N 0.000 description 1
- MHBSUKYVBZVQRW-HJWJTTGWSA-N Pro-Phe-Ile Chemical compound [H]N1CCC[C@H]1C(=O)N[C@@H](CC1=CC=CC=C1)C(=O)N[C@@H]([C@@H](C)CC)C(O)=O MHBSUKYVBZVQRW-HJWJTTGWSA-N 0.000 description 1
- BUEIYHBJHCDAMI-UFYCRDLUSA-N Pro-Phe-Phe Chemical compound [H]N1CCC[C@H]1C(=O)N[C@@H](CC1=CC=CC=C1)C(=O)N[C@@H](CC1=CC=CC=C1)C(O)=O BUEIYHBJHCDAMI-UFYCRDLUSA-N 0.000 description 1
- ZVEQWRWMRFIVSD-HRCADAONSA-N Pro-Phe-Pro Chemical compound C1C[C@H](NC1)C(=O)N[C@@H](CC2=CC=CC=C2)C(=O)N3CCC[C@@H]3C(=O)O ZVEQWRWMRFIVSD-HRCADAONSA-N 0.000 description 1
- XYAFCOJKICBRDU-JYJNAYRXSA-N Pro-Phe-Val Chemical compound [H]N1CCC[C@H]1C(=O)N[C@@H](CC1=CC=CC=C1)C(=O)N[C@@H](C(C)C)C(O)=O XYAFCOJKICBRDU-JYJNAYRXSA-N 0.000 description 1
- RFWXYTJSVDUBBZ-DCAQKATOSA-N Pro-Pro-Glu Chemical compound OC(=O)CC[C@@H](C(O)=O)NC(=O)[C@@H]1CCCN1C(=O)[C@H]1NCCC1 RFWXYTJSVDUBBZ-DCAQKATOSA-N 0.000 description 1
- LEIKGVHQTKHOLM-IUCAKERBSA-N Pro-Pro-Gly Chemical compound OC(=O)CNC(=O)[C@@H]1CCCN1C(=O)[C@H]1NCCC1 LEIKGVHQTKHOLM-IUCAKERBSA-N 0.000 description 1
- SEZGGSHLMROBFX-CIUDSAMLSA-N Pro-Ser-Gln Chemical compound [H]N1CCC[C@H]1C(=O)N[C@@H](CO)C(=O)N[C@@H](CCC(N)=O)C(O)=O SEZGGSHLMROBFX-CIUDSAMLSA-N 0.000 description 1
- GBUNEGKQPSAMNK-QTKMDUPCSA-N Pro-Thr-His Chemical compound C[C@H]([C@@H](C(=O)N[C@@H](CC1=CN=CN1)C(=O)O)NC(=O)[C@@H]2CCCN2)O GBUNEGKQPSAMNK-QTKMDUPCSA-N 0.000 description 1
- QHSSUIHLAIWXEE-IHRRRGAJSA-N Pro-Tyr-Asn Chemical compound [H]N1CCC[C@H]1C(=O)N[C@@H](CC1=CC=C(O)C=C1)C(=O)N[C@@H](CC(N)=O)C(O)=O QHSSUIHLAIWXEE-IHRRRGAJSA-N 0.000 description 1
- QDDJNKWPTJHROJ-UFYCRDLUSA-N Pro-Tyr-Tyr Chemical compound C([C@@H](C(=O)O)NC(=O)[C@H](CC=1C=CC(O)=CC=1)NC(=O)[C@H]1NCCC1)C1=CC=C(O)C=C1 QDDJNKWPTJHROJ-UFYCRDLUSA-N 0.000 description 1
- FHJQROWZEJFZPO-SRVKXCTJSA-N Pro-Val-Val Chemical compound CC(C)[C@@H](C(O)=O)NC(=O)[C@H](C(C)C)NC(=O)[C@@H]1CCCN1 FHJQROWZEJFZPO-SRVKXCTJSA-N 0.000 description 1
- 108010001267 Protein Subunits Proteins 0.000 description 1
- 102000002067 Protein Subunits Human genes 0.000 description 1
- 241000589624 Pseudomonas amygdali pv. tabaci Species 0.000 description 1
- 239000012614 Q-Sepharose Substances 0.000 description 1
- 238000002123 RNA extraction Methods 0.000 description 1
- 108010092799 RNA-directed DNA polymerase Proteins 0.000 description 1
- 241000589771 Ralstonia solanacearum Species 0.000 description 1
- 241000922190 Restio Species 0.000 description 1
- 238000012300 Sequence Analysis Methods 0.000 description 1
- IDQFQFVEWMWRQQ-DLOVCJGASA-N Ser-Ala-Phe Chemical compound [H]N[C@@H](CO)C(=O)N[C@@H](C)C(=O)N[C@@H](CC1=CC=CC=C1)C(O)=O IDQFQFVEWMWRQQ-DLOVCJGASA-N 0.000 description 1
- BRKHVZNDAOMAHX-BIIVOSGPSA-N Ser-Ala-Pro Chemical compound C[C@@H](C(=O)N1CCC[C@@H]1C(=O)O)NC(=O)[C@H](CO)N BRKHVZNDAOMAHX-BIIVOSGPSA-N 0.000 description 1
- YQHZVYJAGWMHES-ZLUOBGJFSA-N Ser-Ala-Ser Chemical compound OC[C@H](N)C(=O)N[C@@H](C)C(=O)N[C@@H](CO)C(O)=O YQHZVYJAGWMHES-ZLUOBGJFSA-N 0.000 description 1
- PZZJMBYSYAKYPK-UWJYBYFXSA-N Ser-Ala-Tyr Chemical compound [H]N[C@@H](CO)C(=O)N[C@@H](C)C(=O)N[C@@H](CC1=CC=C(O)C=C1)C(O)=O PZZJMBYSYAKYPK-UWJYBYFXSA-N 0.000 description 1
- JJKSSJVYOVRJMZ-FXQIFTODSA-N Ser-Arg-Cys Chemical compound C(C[C@@H](C(=O)N[C@@H](CS)C(=O)O)NC(=O)[C@H](CO)N)CN=C(N)N JJKSSJVYOVRJMZ-FXQIFTODSA-N 0.000 description 1
- HQTKVSCNCDLXSX-BQBZGAKWSA-N Ser-Arg-Gly Chemical compound [H]N[C@@H](CO)C(=O)N[C@@H](CCCNC(N)=N)C(=O)NCC(O)=O HQTKVSCNCDLXSX-BQBZGAKWSA-N 0.000 description 1
- QFBNNYNWKYKVJO-DCAQKATOSA-N Ser-Arg-Lys Chemical compound NCCCC[C@@H](C(O)=O)NC(=O)[C@@H](NC(=O)[C@@H](N)CO)CCCN=C(N)N QFBNNYNWKYKVJO-DCAQKATOSA-N 0.000 description 1
- DKKGAAJTDKHWOD-BIIVOSGPSA-N Ser-Asn-Pro Chemical compound C1C[C@@H](N(C1)C(=O)[C@H](CC(=O)N)NC(=O)[C@H](CO)N)C(=O)O DKKGAAJTDKHWOD-BIIVOSGPSA-N 0.000 description 1
- OHKLFYXEOGGGCK-ZLUOBGJFSA-N Ser-Asp-Asn Chemical compound [H]N[C@@H](CO)C(=O)N[C@@H](CC(O)=O)C(=O)N[C@@H](CC(N)=O)C(O)=O OHKLFYXEOGGGCK-ZLUOBGJFSA-N 0.000 description 1
- SFZKGGOGCNQPJY-CIUDSAMLSA-N Ser-Asp-His Chemical compound C1=C(NC=N1)C[C@@H](C(=O)O)NC(=O)[C@H](CC(=O)O)NC(=O)[C@H](CO)N SFZKGGOGCNQPJY-CIUDSAMLSA-N 0.000 description 1
- ULVMNZOKDBHKKI-ACZMJKKPSA-N Ser-Gln-Asp Chemical compound [H]N[C@@H](CO)C(=O)N[C@@H](CCC(N)=O)C(=O)N[C@@H](CC(O)=O)C(O)=O ULVMNZOKDBHKKI-ACZMJKKPSA-N 0.000 description 1
- ZOHGLPQGEHSLPD-FXQIFTODSA-N Ser-Gln-Glu Chemical compound [H]N[C@@H](CO)C(=O)N[C@@H](CCC(N)=O)C(=O)N[C@@H](CCC(O)=O)C(O)=O ZOHGLPQGEHSLPD-FXQIFTODSA-N 0.000 description 1
- YPUSXTWURJANKF-KBIXCLLPSA-N Ser-Gln-Ile Chemical compound [H]N[C@@H](CO)C(=O)N[C@@H](CCC(N)=O)C(=O)N[C@@H]([C@@H](C)CC)C(O)=O YPUSXTWURJANKF-KBIXCLLPSA-N 0.000 description 1
- GWMXFEMMBHOKDX-AVGNSLFASA-N Ser-Gln-Phe Chemical compound OC[C@H](N)C(=O)N[C@@H](CCC(N)=O)C(=O)N[C@H](C(O)=O)CC1=CC=CC=C1 GWMXFEMMBHOKDX-AVGNSLFASA-N 0.000 description 1
- YQQKYAZABFEYAF-FXQIFTODSA-N Ser-Glu-Gln Chemical compound [H]N[C@@H](CO)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CCC(N)=O)C(O)=O YQQKYAZABFEYAF-FXQIFTODSA-N 0.000 description 1
- VQBCMLMPEWPUTB-ACZMJKKPSA-N Ser-Glu-Ser Chemical compound [H]N[C@@H](CO)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CO)C(O)=O VQBCMLMPEWPUTB-ACZMJKKPSA-N 0.000 description 1
- OHKFXGKHSJKKAL-NRPADANISA-N Ser-Glu-Val Chemical compound [H]N[C@@H](CO)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](C(C)C)C(O)=O OHKFXGKHSJKKAL-NRPADANISA-N 0.000 description 1
- MIJWOJAXARLEHA-WDSKDSINSA-N Ser-Gly-Glu Chemical compound OC[C@H](N)C(=O)NCC(=O)N[C@H](C(O)=O)CCC(O)=O MIJWOJAXARLEHA-WDSKDSINSA-N 0.000 description 1
- YMTLKLXDFCSCNX-BYPYZUCNSA-N Ser-Gly-Gly Chemical compound OC[C@H](N)C(=O)NCC(=O)NCC(O)=O YMTLKLXDFCSCNX-BYPYZUCNSA-N 0.000 description 1
- IXCHOHLPHNGFTJ-YUMQZZPRSA-N Ser-Gly-His Chemical compound C1=C(NC=N1)C[C@@H](C(=O)O)NC(=O)CNC(=O)[C@H](CO)N IXCHOHLPHNGFTJ-YUMQZZPRSA-N 0.000 description 1
- SFTZWNJFZYOLBD-ZDLURKLDSA-N Ser-Gly-Thr Chemical compound C[C@@H](O)[C@@H](C(O)=O)NC(=O)CNC(=O)[C@@H](N)CO SFTZWNJFZYOLBD-ZDLURKLDSA-N 0.000 description 1
- XXXAXOWMBOKTRN-XPUUQOCRSA-N Ser-Gly-Val Chemical compound [H]N[C@@H](CO)C(=O)NCC(=O)N[C@@H](C(C)C)C(O)=O XXXAXOWMBOKTRN-XPUUQOCRSA-N 0.000 description 1
- JEHPKECJCALLRW-CUJWVEQBSA-N Ser-His-Thr Chemical compound [H]N[C@@H](CO)C(=O)N[C@@H](CC1=CNC=N1)C(=O)N[C@@H]([C@@H](C)O)C(O)=O JEHPKECJCALLRW-CUJWVEQBSA-N 0.000 description 1
- NNFMANHDYSVNIO-DCAQKATOSA-N Ser-Lys-Arg Chemical compound [H]N[C@@H](CO)C(=O)N[C@@H](CCCCN)C(=O)N[C@@H](CCCNC(N)=N)C(O)=O NNFMANHDYSVNIO-DCAQKATOSA-N 0.000 description 1
- FPCGZYMRFFIYIH-CIUDSAMLSA-N Ser-Lys-Ser Chemical compound [H]N[C@@H](CO)C(=O)N[C@@H](CCCCN)C(=O)N[C@@H](CO)C(O)=O FPCGZYMRFFIYIH-CIUDSAMLSA-N 0.000 description 1
- QJKPECIAWNNKIT-KKUMJFAQSA-N Ser-Lys-Tyr Chemical compound [H]N[C@@H](CO)C(=O)N[C@@H](CCCCN)C(=O)N[C@@H](CC1=CC=C(O)C=C1)C(O)=O QJKPECIAWNNKIT-KKUMJFAQSA-N 0.000 description 1
- NQZFFLBPNDLTPO-DLOVCJGASA-N Ser-Phe-Ala Chemical compound C[C@@H](C(=O)O)NC(=O)[C@H](CC1=CC=CC=C1)NC(=O)[C@H](CO)N NQZFFLBPNDLTPO-DLOVCJGASA-N 0.000 description 1
- JAWGSPUJAXYXJA-IHRRRGAJSA-N Ser-Phe-Arg Chemical compound NC(N)=NCCC[C@@H](C(O)=O)NC(=O)[C@@H](NC(=O)[C@H](CO)N)CC1=CC=CC=C1 JAWGSPUJAXYXJA-IHRRRGAJSA-N 0.000 description 1
- FZEUTKVQGMVGHW-AVGNSLFASA-N Ser-Phe-Gln Chemical compound [H]N[C@@H](CO)C(=O)N[C@@H](CC1=CC=CC=C1)C(=O)N[C@@H](CCC(N)=O)C(O)=O FZEUTKVQGMVGHW-AVGNSLFASA-N 0.000 description 1
- UPLYXVPQLJVWMM-KKUMJFAQSA-N Ser-Phe-Leu Chemical compound [H]N[C@@H](CO)C(=O)N[C@@H](CC1=CC=CC=C1)C(=O)N[C@@H](CC(C)C)C(O)=O UPLYXVPQLJVWMM-KKUMJFAQSA-N 0.000 description 1
- XVWDJUROVRQKAE-KKUMJFAQSA-N Ser-Phe-Lys Chemical compound NCCCC[C@@H](C(O)=O)NC(=O)[C@@H](NC(=O)[C@@H](N)CO)CC1=CC=CC=C1 XVWDJUROVRQKAE-KKUMJFAQSA-N 0.000 description 1
- RRVFEDGUXSYWOW-BZSNNMDCSA-N Ser-Phe-Phe Chemical compound [H]N[C@@H](CO)C(=O)N[C@@H](CC1=CC=CC=C1)C(=O)N[C@@H](CC1=CC=CC=C1)C(O)=O RRVFEDGUXSYWOW-BZSNNMDCSA-N 0.000 description 1
- MQUZANJDFOQOBX-SRVKXCTJSA-N Ser-Phe-Ser Chemical compound [H]N[C@@H](CO)C(=O)N[C@@H](CC1=CC=CC=C1)C(=O)N[C@@H](CO)C(O)=O MQUZANJDFOQOBX-SRVKXCTJSA-N 0.000 description 1
- BSXKBOUZDAZXHE-CIUDSAMLSA-N Ser-Pro-Glu Chemical compound [H]N[C@@H](CO)C(=O)N1CCC[C@H]1C(=O)N[C@@H](CCC(O)=O)C(O)=O BSXKBOUZDAZXHE-CIUDSAMLSA-N 0.000 description 1
- DINQYZRMXGWWTG-GUBZILKMSA-N Ser-Pro-Pro Chemical compound OC[C@H](N)C(=O)N1CCC[C@H]1C(=O)N1[C@H](C(O)=O)CCC1 DINQYZRMXGWWTG-GUBZILKMSA-N 0.000 description 1
- AZWNCEBQZXELEZ-FXQIFTODSA-N Ser-Pro-Ser Chemical compound OC[C@H](N)C(=O)N1CCC[C@H]1C(=O)N[C@@H](CO)C(O)=O AZWNCEBQZXELEZ-FXQIFTODSA-N 0.000 description 1
- CKDXFSPMIDSMGV-GUBZILKMSA-N Ser-Pro-Val Chemical compound [H]N[C@@H](CO)C(=O)N1CCC[C@H]1C(=O)N[C@@H](C(C)C)C(O)=O CKDXFSPMIDSMGV-GUBZILKMSA-N 0.000 description 1
- KQNDIKOYWZTZIX-FXQIFTODSA-N Ser-Ser-Arg Chemical compound OC[C@H](N)C(=O)N[C@@H](CO)C(=O)N[C@H](C(O)=O)CCCNC(N)=N KQNDIKOYWZTZIX-FXQIFTODSA-N 0.000 description 1
- FZXOPYUEQGDGMS-ACZMJKKPSA-N Ser-Ser-Gln Chemical compound [H]N[C@@H](CO)C(=O)N[C@@H](CO)C(=O)N[C@@H](CCC(N)=O)C(O)=O FZXOPYUEQGDGMS-ACZMJKKPSA-N 0.000 description 1
- SRSPTFBENMJHMR-WHFBIAKZSA-N Ser-Ser-Gly Chemical compound OC[C@H](N)C(=O)N[C@@H](CO)C(=O)NCC(O)=O SRSPTFBENMJHMR-WHFBIAKZSA-N 0.000 description 1
- SDFUZKIAHWRUCS-QEJZJMRPSA-N Ser-Trp-Glu Chemical compound C1=CC=C2C(=C1)C(=CN2)C[C@@H](C(=O)N[C@@H](CCC(=O)O)C(=O)O)NC(=O)[C@H](CO)N SDFUZKIAHWRUCS-QEJZJMRPSA-N 0.000 description 1
- FGBLCMLXHRPVOF-IHRRRGAJSA-N Ser-Tyr-Arg Chemical compound [H]N[C@@H](CO)C(=O)N[C@@H](CC1=CC=C(O)C=C1)C(=O)N[C@@H](CCCNC(N)=N)C(O)=O FGBLCMLXHRPVOF-IHRRRGAJSA-N 0.000 description 1
- PQEQXWRVHQAAKS-SRVKXCTJSA-N Ser-Tyr-Asn Chemical compound NC(=O)C[C@@H](C(O)=O)NC(=O)[C@@H](NC(=O)[C@H](CO)N)CC1=CC=C(O)C=C1 PQEQXWRVHQAAKS-SRVKXCTJSA-N 0.000 description 1
- UBTNVMGPMYDYIU-HJPIBITLSA-N Ser-Tyr-Ile Chemical compound [H]N[C@@H](CO)C(=O)N[C@@H](CC1=CC=C(O)C=C1)C(=O)N[C@@H]([C@@H](C)CC)C(O)=O UBTNVMGPMYDYIU-HJPIBITLSA-N 0.000 description 1
- PLQWGQUNUPMNOD-KKUMJFAQSA-N Ser-Tyr-Leu Chemical compound [H]N[C@@H](CO)C(=O)N[C@@H](CC1=CC=C(O)C=C1)C(=O)N[C@@H](CC(C)C)C(O)=O PLQWGQUNUPMNOD-KKUMJFAQSA-N 0.000 description 1
- OQSQCUWQOIHECT-YJRXYDGGSA-N Ser-Tyr-Thr Chemical compound [H]N[C@@H](CO)C(=O)N[C@@H](CC1=CC=C(O)C=C1)C(=O)N[C@@H]([C@@H](C)O)C(O)=O OQSQCUWQOIHECT-YJRXYDGGSA-N 0.000 description 1
- 240000003829 Sorghum propinquum Species 0.000 description 1
- 241000534520 Stenocarpus Species 0.000 description 1
- 102100021588 Sterol carrier protein 2 Human genes 0.000 description 1
- QAOWNCQODCNURD-UHFFFAOYSA-N Sulfuric acid Chemical compound OS(O)(=O)=O QAOWNCQODCNURD-UHFFFAOYSA-N 0.000 description 1
- UZMAPBJVXOGOFT-UHFFFAOYSA-N Syringetin Natural products COC1=C(O)C(OC)=CC(C2=C(C(=O)C3=C(O)C=C(O)C=C3O2)O)=C1 UZMAPBJVXOGOFT-UHFFFAOYSA-N 0.000 description 1
- 241000867810 Tetragona Species 0.000 description 1
- 101710203193 Thaumatin-like protein Proteins 0.000 description 1
- 235000005764 Theobroma cacao ssp. cacao Nutrition 0.000 description 1
- 235000005767 Theobroma cacao ssp. sphaerocarpum Nutrition 0.000 description 1
- 241000561282 Thielaviopsis basicola Species 0.000 description 1
- 108010076830 Thionins Proteins 0.000 description 1
- PXQUBKWZENPDGE-CIQUZCHMSA-N Thr-Ala-Ile Chemical compound CC[C@H](C)[C@@H](C(=O)O)NC(=O)[C@H](C)NC(=O)[C@H]([C@@H](C)O)N PXQUBKWZENPDGE-CIQUZCHMSA-N 0.000 description 1
- UKBSDLHIKIXJKH-HJGDQZAQSA-N Thr-Arg-Glu Chemical compound [H]N[C@@H]([C@@H](C)O)C(=O)N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CCC(O)=O)C(O)=O UKBSDLHIKIXJKH-HJGDQZAQSA-N 0.000 description 1
- TWLMXDWFVNEFFK-FJXKBIBVSA-N Thr-Arg-Gly Chemical compound [H]N[C@@H]([C@@H](C)O)C(=O)N[C@@H](CCCNC(N)=N)C(=O)NCC(O)=O TWLMXDWFVNEFFK-FJXKBIBVSA-N 0.000 description 1
- UTSWGQNAQRIHAI-UNQGMJICSA-N Thr-Arg-Phe Chemical compound NC(N)=NCCC[C@H](NC(=O)[C@@H](N)[C@H](O)C)C(=O)N[C@H](C(O)=O)CC1=CC=CC=C1 UTSWGQNAQRIHAI-UNQGMJICSA-N 0.000 description 1
- TZKPNGDGUVREEB-FOHZUACHSA-N Thr-Asn-Gly Chemical compound C[C@@H](O)[C@H](N)C(=O)N[C@@H](CC(N)=O)C(=O)NCC(O)=O TZKPNGDGUVREEB-FOHZUACHSA-N 0.000 description 1
- PZVGOVRNGKEFCB-KKHAAJSZSA-N Thr-Asn-Val Chemical compound C[C@H]([C@@H](C(=O)N[C@@H](CC(=O)N)C(=O)N[C@@H](C(C)C)C(=O)O)N)O PZVGOVRNGKEFCB-KKHAAJSZSA-N 0.000 description 1
- DSLHSTIUAPKERR-XGEHTFHBSA-N Thr-Cys-Val Chemical compound [H]N[C@@H]([C@@H](C)O)C(=O)N[C@@H](CS)C(=O)N[C@@H](C(C)C)C(O)=O DSLHSTIUAPKERR-XGEHTFHBSA-N 0.000 description 1
- VGYBYGQXZJDZJU-XQXXSGGOSA-N Thr-Glu-Ala Chemical compound [H]N[C@@H]([C@@H](C)O)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](C)C(O)=O VGYBYGQXZJDZJU-XQXXSGGOSA-N 0.000 description 1
- VOHWDZNIESHTFW-XKBZYTNZSA-N Thr-Glu-Cys Chemical compound C[C@H]([C@@H](C(=O)N[C@@H](CCC(=O)O)C(=O)N[C@@H](CS)C(=O)O)N)O VOHWDZNIESHTFW-XKBZYTNZSA-N 0.000 description 1
- UDQBCBUXAQIZAK-GLLZPBPUSA-N Thr-Glu-Glu Chemical compound [H]N[C@@H]([C@@H](C)O)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CCC(O)=O)C(O)=O UDQBCBUXAQIZAK-GLLZPBPUSA-N 0.000 description 1
- SLUWOCTZVGMURC-BFHQHQDPSA-N Thr-Gly-Ala Chemical compound C[C@@H](O)[C@H](N)C(=O)NCC(=O)N[C@@H](C)C(O)=O SLUWOCTZVGMURC-BFHQHQDPSA-N 0.000 description 1
- XTCNBOBTROGWMW-RWRJDSDZSA-N Thr-Ile-Glu Chemical compound CC[C@H](C)[C@@H](C(=O)N[C@@H](CCC(=O)O)C(=O)O)NC(=O)[C@H]([C@@H](C)O)N XTCNBOBTROGWMW-RWRJDSDZSA-N 0.000 description 1
- AHOLTQCAVBSUDP-PPCPHDFISA-N Thr-Ile-Lys Chemical compound CC[C@H](C)[C@H](NC(=O)[C@@H](N)[C@@H](C)O)C(=O)N[C@@H](CCCCN)C(O)=O AHOLTQCAVBSUDP-PPCPHDFISA-N 0.000 description 1
- VTVVYQOXJCZVEB-WDCWCFNPSA-N Thr-Leu-Glu Chemical compound [H]N[C@@H]([C@@H](C)O)C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](CCC(O)=O)C(O)=O VTVVYQOXJCZVEB-WDCWCFNPSA-N 0.000 description 1
- MEJHFIOYJHTWMK-VOAKCMCISA-N Thr-Leu-Leu Chemical compound CC(C)C[C@@H](C(O)=O)NC(=O)[C@H](CC(C)C)NC(=O)[C@@H](N)[C@@H](C)O MEJHFIOYJHTWMK-VOAKCMCISA-N 0.000 description 1
- FIFDDJFLNVAVMS-RHYQMDGZSA-N Thr-Leu-Met Chemical compound [H]N[C@@H]([C@@H](C)O)C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](CCSC)C(O)=O FIFDDJFLNVAVMS-RHYQMDGZSA-N 0.000 description 1
- PRNGXSILMXSWQQ-OEAJRASXSA-N Thr-Leu-Phe Chemical compound [H]N[C@@H]([C@@H](C)O)C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](CC1=CC=CC=C1)C(O)=O PRNGXSILMXSWQQ-OEAJRASXSA-N 0.000 description 1
- SPVHQURZJCUDQC-VOAKCMCISA-N Thr-Lys-Leu Chemical compound [H]N[C@@H]([C@@H](C)O)C(=O)N[C@@H](CCCCN)C(=O)N[C@@H](CC(C)C)C(O)=O SPVHQURZJCUDQC-VOAKCMCISA-N 0.000 description 1
- KZURUCDWKDEAFZ-XVSYOHENSA-N Thr-Phe-Asn Chemical compound C[C@H]([C@@H](C(=O)N[C@@H](CC1=CC=CC=C1)C(=O)N[C@@H](CC(=O)N)C(=O)O)N)O KZURUCDWKDEAFZ-XVSYOHENSA-N 0.000 description 1
- HSQXHRIRJSFDOH-URLPEUOOSA-N Thr-Phe-Ile Chemical compound [H]N[C@@H]([C@@H](C)O)C(=O)N[C@@H](CC1=CC=CC=C1)C(=O)N[C@@H]([C@@H](C)CC)C(O)=O HSQXHRIRJSFDOH-URLPEUOOSA-N 0.000 description 1
- WNQJTLATMXYSEL-OEAJRASXSA-N Thr-Phe-Leu Chemical compound [H]N[C@@H]([C@@H](C)O)C(=O)N[C@@H](CC1=CC=CC=C1)C(=O)N[C@@H](CC(C)C)C(O)=O WNQJTLATMXYSEL-OEAJRASXSA-N 0.000 description 1
- VEIKMWOMUYMMMK-FCLVOEFKSA-N Thr-Phe-Phe Chemical compound C([C@H](NC(=O)[C@@H](N)[C@H](O)C)C(=O)N[C@@H](CC=1C=CC=CC=1)C(O)=O)C1=CC=CC=C1 VEIKMWOMUYMMMK-FCLVOEFKSA-N 0.000 description 1
- NWECYMJLJGCBOD-UNQGMJICSA-N Thr-Phe-Val Chemical compound [H]N[C@@H]([C@@H](C)O)C(=O)N[C@@H](CC1=CC=CC=C1)C(=O)N[C@@H](C(C)C)C(O)=O NWECYMJLJGCBOD-UNQGMJICSA-N 0.000 description 1
- MUAFDCVOHYAFNG-RCWTZXSCSA-N Thr-Pro-Arg Chemical compound [H]N[C@@H]([C@@H](C)O)C(=O)N1CCC[C@H]1C(=O)N[C@@H](CCCNC(N)=N)C(O)=O MUAFDCVOHYAFNG-RCWTZXSCSA-N 0.000 description 1
- XKWABWFMQXMUMT-HJGDQZAQSA-N Thr-Pro-Glu Chemical compound [H]N[C@@H]([C@@H](C)O)C(=O)N1CCC[C@H]1C(=O)N[C@@H](CCC(O)=O)C(O)=O XKWABWFMQXMUMT-HJGDQZAQSA-N 0.000 description 1
- STUAPCLEDMKXKL-LKXGYXEUSA-N Thr-Ser-Asn Chemical compound [H]N[C@@H]([C@@H](C)O)C(=O)N[C@@H](CO)C(=O)N[C@@H](CC(N)=O)C(O)=O STUAPCLEDMKXKL-LKXGYXEUSA-N 0.000 description 1
- SGAOHNPSEPVAFP-ZDLURKLDSA-N Thr-Ser-Gly Chemical compound [H]N[C@@H]([C@@H](C)O)C(=O)N[C@@H](CO)C(=O)NCC(O)=O SGAOHNPSEPVAFP-ZDLURKLDSA-N 0.000 description 1
- AHERARIZBPOMNU-KATARQTJSA-N Thr-Ser-Leu Chemical compound [H]N[C@@H]([C@@H](C)O)C(=O)N[C@@H](CO)C(=O)N[C@@H](CC(C)C)C(O)=O AHERARIZBPOMNU-KATARQTJSA-N 0.000 description 1
- YOPQYBJJNSIQGZ-JNPHEJMOSA-N Thr-Tyr-Tyr Chemical compound C([C@H](NC(=O)[C@@H](N)[C@H](O)C)C(=O)N[C@@H](CC=1C=CC(O)=CC=1)C(O)=O)C1=CC=C(O)C=C1 YOPQYBJJNSIQGZ-JNPHEJMOSA-N 0.000 description 1
- PWONLXBUSVIZPH-RHYQMDGZSA-N Thr-Val-Lys Chemical compound C[C@H]([C@@H](C(=O)N[C@@H](C(C)C)C(=O)N[C@@H](CCCCN)C(=O)O)N)O PWONLXBUSVIZPH-RHYQMDGZSA-N 0.000 description 1
- BPGDJSUFQKWUBK-KJEVXHAQSA-N Thr-Val-Tyr Chemical compound C[C@@H](O)[C@H](N)C(=O)N[C@@H](C(C)C)C(=O)N[C@H](C(O)=O)CC1=CC=C(O)C=C1 BPGDJSUFQKWUBK-KJEVXHAQSA-N 0.000 description 1
- 241000723873 Tobacco mosaic virus Species 0.000 description 1
- 239000007997 Tricine buffer Substances 0.000 description 1
- DTQVDTLACAAQTR-UHFFFAOYSA-M Trifluoroacetate Chemical compound [O-]C(=O)C(F)(F)F DTQVDTLACAAQTR-UHFFFAOYSA-M 0.000 description 1
- 239000007983 Tris buffer Substances 0.000 description 1
- SCQBNMKLZVCXNX-ZFWWWQNUSA-N Trp-Arg-Gly Chemical compound C1=CC=C2C(=C1)C(=CN2)C[C@@H](C(=O)N[C@@H](CCCN=C(N)N)C(=O)NCC(=O)O)N SCQBNMKLZVCXNX-ZFWWWQNUSA-N 0.000 description 1
- LTLBNCDNXQCOLB-UBHSHLNASA-N Trp-Asp-Ser Chemical compound C1=CC=C2C(C[C@H](N)C(=O)N[C@@H](CC(O)=O)C(=O)N[C@@H](CO)C(O)=O)=CNC2=C1 LTLBNCDNXQCOLB-UBHSHLNASA-N 0.000 description 1
- SNJAPSVIPKUMCK-NWLDYVSISA-N Trp-Glu-Thr Chemical compound [H]N[C@@H](CC1=CNC2=C1C=CC=C2)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H]([C@@H](C)O)C(O)=O SNJAPSVIPKUMCK-NWLDYVSISA-N 0.000 description 1
- PVRRBEROBJQPJX-SZMVWBNQSA-N Trp-His-Gln Chemical compound C1=CC=C2C(=C1)C(=CN2)C[C@@H](C(=O)N[C@@H](CC3=CN=CN3)C(=O)N[C@@H](CCC(=O)N)C(=O)O)N PVRRBEROBJQPJX-SZMVWBNQSA-N 0.000 description 1
- UHXOYRWHIQZAKV-SZMVWBNQSA-N Trp-Pro-Arg Chemical compound O=C([C@H](CC=1C2=CC=CC=C2NC=1)N)N1CCC[C@H]1C(=O)N[C@@H](CCCN=C(N)N)C(O)=O UHXOYRWHIQZAKV-SZMVWBNQSA-N 0.000 description 1
- AKFLVKKWVZMFOT-IHRRRGAJSA-N Tyr-Arg-Asn Chemical compound [H]N[C@@H](CC1=CC=C(O)C=C1)C(=O)N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CC(N)=O)C(O)=O AKFLVKKWVZMFOT-IHRRRGAJSA-N 0.000 description 1
- GFZQWWDXJVGEMW-ULQDDVLXSA-N Tyr-Arg-Lys Chemical compound C1=CC(=CC=C1C[C@@H](C(=O)N[C@@H](CCCN=C(N)N)C(=O)N[C@@H](CCCCN)C(=O)O)N)O GFZQWWDXJVGEMW-ULQDDVLXSA-N 0.000 description 1
- QYSBJAUCUKHSLU-JYJNAYRXSA-N Tyr-Arg-Val Chemical compound [H]N[C@@H](CC1=CC=C(O)C=C1)C(=O)N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](C(C)C)C(O)=O QYSBJAUCUKHSLU-JYJNAYRXSA-N 0.000 description 1
- MTEQZJFSEMXXRK-CFMVVWHZSA-N Tyr-Asn-Ile Chemical compound CC[C@H](C)[C@@H](C(=O)O)NC(=O)[C@H](CC(=O)N)NC(=O)[C@H](CC1=CC=C(C=C1)O)N MTEQZJFSEMXXRK-CFMVVWHZSA-N 0.000 description 1
- NRFTYDWKWGJLAR-MELADBBJSA-N Tyr-Asp-Pro Chemical compound C1C[C@@H](N(C1)C(=O)[C@H](CC(=O)O)NC(=O)[C@H](CC2=CC=C(C=C2)O)N)C(=O)O NRFTYDWKWGJLAR-MELADBBJSA-N 0.000 description 1
- NZFCWALTLNFHHC-JYJNAYRXSA-N Tyr-Glu-Leu Chemical compound CC(C)C[C@@H](C(O)=O)NC(=O)[C@H](CCC(O)=O)NC(=O)[C@@H](N)CC1=CC=C(O)C=C1 NZFCWALTLNFHHC-JYJNAYRXSA-N 0.000 description 1
- KOVXHANYYYMBRF-IRIUXVKKSA-N Tyr-Glu-Thr Chemical compound C[C@H]([C@@H](C(=O)O)NC(=O)[C@H](CCC(=O)O)NC(=O)[C@H](CC1=CC=C(C=C1)O)N)O KOVXHANYYYMBRF-IRIUXVKKSA-N 0.000 description 1
- CDHQEOXPWBDFPL-QWRGUYRKSA-N Tyr-Gly-Asn Chemical compound NC(=O)C[C@@H](C(O)=O)NC(=O)CNC(=O)[C@@H](N)CC1=CC=C(O)C=C1 CDHQEOXPWBDFPL-QWRGUYRKSA-N 0.000 description 1
- NMKJPMCEKQHRPD-IRXDYDNUSA-N Tyr-Gly-Tyr Chemical compound C([C@H](N)C(=O)NCC(=O)N[C@@H](CC=1C=CC(O)=CC=1)C(O)=O)C1=CC=C(O)C=C1 NMKJPMCEKQHRPD-IRXDYDNUSA-N 0.000 description 1
- GULIUBBXCYPDJU-CQDKDKBSSA-N Tyr-Leu-Ala Chemical compound [O-]C(=O)[C@H](C)NC(=O)[C@H](CC(C)C)NC(=O)[C@@H]([NH3+])CC1=CC=C(O)C=C1 GULIUBBXCYPDJU-CQDKDKBSSA-N 0.000 description 1
- NKUGCYDFQKFVOJ-JYJNAYRXSA-N Tyr-Leu-Gln Chemical compound NC(=O)CC[C@@H](C(O)=O)NC(=O)[C@H](CC(C)C)NC(=O)[C@@H](N)CC1=CC=C(O)C=C1 NKUGCYDFQKFVOJ-JYJNAYRXSA-N 0.000 description 1
- FGVFBDZSGQTYQX-UFYCRDLUSA-N Tyr-Phe-Val Chemical compound [H]N[C@@H](CC1=CC=C(O)C=C1)C(=O)N[C@@H](CC1=CC=CC=C1)C(=O)N[C@@H](C(C)C)C(O)=O FGVFBDZSGQTYQX-UFYCRDLUSA-N 0.000 description 1
- BCOBSVIZMQXKFY-KKUMJFAQSA-N Tyr-Ser-His Chemical compound C1=CC(=CC=C1C[C@@H](C(=O)N[C@@H](CO)C(=O)N[C@@H](CC2=CN=CN2)C(=O)O)N)O BCOBSVIZMQXKFY-KKUMJFAQSA-N 0.000 description 1
- MQGGXGKQSVEQHR-KKUMJFAQSA-N Tyr-Ser-Leu Chemical compound CC(C)C[C@@H](C(O)=O)NC(=O)[C@H](CO)NC(=O)[C@@H](N)CC1=CC=C(O)C=C1 MQGGXGKQSVEQHR-KKUMJFAQSA-N 0.000 description 1
- WYOBRXPIZVKNMF-IRXDYDNUSA-N Tyr-Tyr-Gly Chemical compound C([C@H](N)C(=O)N[C@@H](CC=1C=CC(O)=CC=1)C(=O)NCC(O)=O)C1=CC=C(O)C=C1 WYOBRXPIZVKNMF-IRXDYDNUSA-N 0.000 description 1
- ZLFHAAGHGQBQQN-AEJSXWLSSA-N Val-Ala-Pro Chemical compound C[C@@H](C(=O)N1CCC[C@@H]1C(=O)O)NC(=O)[C@H](C(C)C)N ZLFHAAGHGQBQQN-AEJSXWLSSA-N 0.000 description 1
- UUYCNAXCCDNULB-QXEWZRGKSA-N Val-Arg-Asn Chemical compound CC(C)[C@H](N)C(=O)N[C@@H](CCCN=C(N)N)C(=O)N[C@@H](CC(N)=O)C(O)=O UUYCNAXCCDNULB-QXEWZRGKSA-N 0.000 description 1
- JIODCDXKCJRMEH-NHCYSSNCSA-N Val-Arg-Gln Chemical compound CC(C)[C@@H](C(=O)N[C@@H](CCCN=C(N)N)C(=O)N[C@@H](CCC(=O)N)C(=O)O)N JIODCDXKCJRMEH-NHCYSSNCSA-N 0.000 description 1
- KKHRWGYHBZORMQ-NHCYSSNCSA-N Val-Arg-Glu Chemical compound CC(C)[C@@H](C(=O)N[C@@H](CCCN=C(N)N)C(=O)N[C@@H](CCC(=O)O)C(=O)O)N KKHRWGYHBZORMQ-NHCYSSNCSA-N 0.000 description 1
- PAPWZOJOLKZEFR-AVGNSLFASA-N Val-Arg-Lys Chemical compound CC(C)[C@@H](C(=O)N[C@@H](CCCN=C(N)N)C(=O)N[C@@H](CCCCN)C(=O)O)N PAPWZOJOLKZEFR-AVGNSLFASA-N 0.000 description 1
- QPZMOUMNTGTEFR-ZKWXMUAHSA-N Val-Asn-Ala Chemical compound C[C@@H](C(=O)O)NC(=O)[C@H](CC(=O)N)NC(=O)[C@H](C(C)C)N QPZMOUMNTGTEFR-ZKWXMUAHSA-N 0.000 description 1
- BYOHPUZJVXWHAE-BYULHYEWSA-N Val-Asn-Asn Chemical compound CC(C)[C@@H](C(=O)N[C@@H](CC(=O)N)C(=O)N[C@@H](CC(=O)N)C(=O)O)N BYOHPUZJVXWHAE-BYULHYEWSA-N 0.000 description 1
- OGNMURQZFMHFFD-NHCYSSNCSA-N Val-Asn-Lys Chemical compound CC(C)[C@@H](C(=O)N[C@@H](CC(=O)N)C(=O)N[C@@H](CCCCN)C(=O)O)N OGNMURQZFMHFFD-NHCYSSNCSA-N 0.000 description 1
- JLFKWDAZBRYCGX-ZKWXMUAHSA-N Val-Asn-Ser Chemical compound CC(C)[C@@H](C(=O)N[C@@H](CC(=O)N)C(=O)N[C@@H](CO)C(=O)O)N JLFKWDAZBRYCGX-ZKWXMUAHSA-N 0.000 description 1
- PFMAFMPJJSHNDW-ZKWXMUAHSA-N Val-Cys-Asn Chemical compound CC(C)[C@@H](C(=O)N[C@@H](CS)C(=O)N[C@@H](CC(=O)N)C(=O)O)N PFMAFMPJJSHNDW-ZKWXMUAHSA-N 0.000 description 1
- DLYOEFGPYTZVSP-AEJSXWLSSA-N Val-Cys-Pro Chemical compound CC(C)[C@@H](C(=O)N[C@@H](CS)C(=O)N1CCC[C@@H]1C(=O)O)N DLYOEFGPYTZVSP-AEJSXWLSSA-N 0.000 description 1
- SZTTYWIUCGSURQ-AUTRQRHGSA-N Val-Glu-Glu Chemical compound CC(C)[C@H](N)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CCC(O)=O)C(O)=O SZTTYWIUCGSURQ-AUTRQRHGSA-N 0.000 description 1
- JPPXDMBGXJBTIB-ULQDDVLXSA-N Val-His-Tyr Chemical compound CC(C)[C@@H](C(=O)N[C@@H](CC1=CN=CN1)C(=O)N[C@@H](CC2=CC=C(C=C2)O)C(=O)O)N JPPXDMBGXJBTIB-ULQDDVLXSA-N 0.000 description 1
- CPGJELLYDQEDRK-NAKRPEOUSA-N Val-Ile-Ala Chemical compound CC[C@H](C)[C@H](NC(=O)[C@@H](N)C(C)C)C(=O)N[C@@H](C)C(O)=O CPGJELLYDQEDRK-NAKRPEOUSA-N 0.000 description 1
- LKUDRJSNRWVGMS-QSFUFRPTSA-N Val-Ile-Asp Chemical compound CC[C@H](C)[C@@H](C(=O)N[C@@H](CC(=O)O)C(=O)O)NC(=O)[C@H](C(C)C)N LKUDRJSNRWVGMS-QSFUFRPTSA-N 0.000 description 1
- WNZSAUMKZQXHNC-UKJIMTQDSA-N Val-Ile-Gln Chemical compound CC[C@H](C)[C@@H](C(=O)N[C@@H](CCC(=O)N)C(=O)O)NC(=O)[C@H](C(C)C)N WNZSAUMKZQXHNC-UKJIMTQDSA-N 0.000 description 1
- FTKXYXACXYOHND-XUXIUFHCSA-N Val-Ile-Leu Chemical compound CC(C)[C@H](N)C(=O)N[C@@H]([C@@H](C)CC)C(=O)N[C@@H](CC(C)C)C(O)=O FTKXYXACXYOHND-XUXIUFHCSA-N 0.000 description 1
- APQIVBCUIUDSMB-OSUNSFLBSA-N Val-Ile-Thr Chemical compound CC[C@H](C)[C@@H](C(=O)N[C@@H]([C@@H](C)O)C(=O)O)NC(=O)[C@H](C(C)C)N APQIVBCUIUDSMB-OSUNSFLBSA-N 0.000 description 1
- DJQIUOKSNRBTSV-CYDGBPFRSA-N Val-Ile-Val Chemical compound CC[C@H](C)[C@@H](C(=O)N[C@@H](C(C)C)C(=O)O)NC(=O)[C@H](C(C)C)N DJQIUOKSNRBTSV-CYDGBPFRSA-N 0.000 description 1
- UMPVMAYCLYMYGA-ONGXEEELSA-N Val-Leu-Gly Chemical compound CC(C)[C@H](N)C(=O)N[C@@H](CC(C)C)C(=O)NCC(O)=O UMPVMAYCLYMYGA-ONGXEEELSA-N 0.000 description 1
- DAVNYIUELQBTAP-XUXIUFHCSA-N Val-Leu-Ile Chemical compound CC[C@H](C)[C@@H](C(=O)O)NC(=O)[C@H](CC(C)C)NC(=O)[C@H](C(C)C)N DAVNYIUELQBTAP-XUXIUFHCSA-N 0.000 description 1
- VPGCVZRRBYOGCD-AVGNSLFASA-N Val-Lys-Val Chemical compound CC(C)[C@H](N)C(=O)N[C@@H](CCCCN)C(=O)N[C@@H](C(C)C)C(O)=O VPGCVZRRBYOGCD-AVGNSLFASA-N 0.000 description 1
- FMQGYTMERWBMSI-HJWJTTGWSA-N Val-Phe-Ile Chemical compound CC[C@H](C)[C@@H](C(=O)O)NC(=O)[C@H](CC1=CC=CC=C1)NC(=O)[C@H](C(C)C)N FMQGYTMERWBMSI-HJWJTTGWSA-N 0.000 description 1
- HJSLDXZAZGFPDK-ULQDDVLXSA-N Val-Phe-Leu Chemical compound CC(C)C[C@@H](C(=O)O)NC(=O)[C@H](CC1=CC=CC=C1)NC(=O)[C@H](C(C)C)N HJSLDXZAZGFPDK-ULQDDVLXSA-N 0.000 description 1
- SJRUJQFQVLMZFW-WPRPVWTQSA-N Val-Pro-Gly Chemical compound CC(C)[C@H](N)C(=O)N1CCC[C@H]1C(=O)NCC(O)=O SJRUJQFQVLMZFW-WPRPVWTQSA-N 0.000 description 1
- DOFAQXCYFQKSHT-SRVKXCTJSA-N Val-Pro-Pro Chemical compound CC(C)[C@H](N)C(=O)N1CCC[C@H]1C(=O)N1[C@H](C(O)=O)CCC1 DOFAQXCYFQKSHT-SRVKXCTJSA-N 0.000 description 1
- QSPOLEBZTMESFY-SRVKXCTJSA-N Val-Pro-Val Chemical compound CC(C)[C@H](N)C(=O)N1CCC[C@H]1C(=O)N[C@@H](C(C)C)C(O)=O QSPOLEBZTMESFY-SRVKXCTJSA-N 0.000 description 1
- GBIUHAYJGWVNLN-AEJSXWLSSA-N Val-Ser-Pro Chemical compound CC(C)[C@@H](C(=O)N[C@@H](CO)C(=O)N1CCC[C@@H]1C(=O)O)N GBIUHAYJGWVNLN-AEJSXWLSSA-N 0.000 description 1
- UVHFONIHVHLDDQ-IFFSRLJSSA-N Val-Thr-Glu Chemical compound C[C@H]([C@@H](C(=O)N[C@@H](CCC(=O)O)C(=O)O)NC(=O)[C@H](C(C)C)N)O UVHFONIHVHLDDQ-IFFSRLJSSA-N 0.000 description 1
- TVGWMCTYUFBXAP-QTKMDUPCSA-N Val-Thr-His Chemical compound C[C@H]([C@@H](C(=O)N[C@@H](CC1=CN=CN1)C(=O)O)NC(=O)[C@H](C(C)C)N)O TVGWMCTYUFBXAP-QTKMDUPCSA-N 0.000 description 1
- PDDJTOSAVNRJRH-UNQGMJICSA-N Val-Thr-Phe Chemical compound C[C@H]([C@@H](C(=O)N[C@@H](CC1=CC=CC=C1)C(=O)O)NC(=O)[C@H](C(C)C)N)O PDDJTOSAVNRJRH-UNQGMJICSA-N 0.000 description 1
- ZLMFVXMJFIWIRE-FHWLQOOXSA-N Val-Trp-Leu Chemical compound CC(C)C[C@@H](C(=O)O)NC(=O)[C@H](CC1=CNC2=CC=CC=C21)NC(=O)[C@H](C(C)C)N ZLMFVXMJFIWIRE-FHWLQOOXSA-N 0.000 description 1
- JXWGBRRVTRAZQA-ULQDDVLXSA-N Val-Tyr-Leu Chemical compound CC(C)C[C@@H](C(=O)O)NC(=O)[C@H](CC1=CC=C(C=C1)O)NC(=O)[C@H](C(C)C)N JXWGBRRVTRAZQA-ULQDDVLXSA-N 0.000 description 1
- ZHWZDZFWBXWPDW-GUBZILKMSA-N Val-Val-Cys Chemical compound CC(C)[C@H](N)C(=O)N[C@@H](C(C)C)C(=O)N[C@@H](CS)C(O)=O ZHWZDZFWBXWPDW-GUBZILKMSA-N 0.000 description 1
- AOILQMZPNLUXCM-AVGNSLFASA-N Val-Val-Lys Chemical compound CC(C)[C@H](N)C(=O)N[C@@H](C(C)C)C(=O)N[C@H](C(O)=O)CCCCN AOILQMZPNLUXCM-AVGNSLFASA-N 0.000 description 1
- JSOXWWFKRJKTMT-WOPDTQHZSA-N Val-Val-Pro Chemical compound CC(C)[C@@H](C(=O)N[C@@H](C(C)C)C(=O)N1CCC[C@@H]1C(=O)O)N JSOXWWFKRJKTMT-WOPDTQHZSA-N 0.000 description 1
- LLJLBRRXKZTTRD-GUBZILKMSA-N Val-Val-Ser Chemical compound CC(C)[C@@H](C(=O)N[C@@H](C(C)C)C(=O)N[C@@H](CO)C(=O)O)N LLJLBRRXKZTTRD-GUBZILKMSA-N 0.000 description 1
- JVGDAEKKZKKZFO-RCWTZXSCSA-N Val-Val-Thr Chemical compound C[C@H]([C@@H](C(=O)O)NC(=O)[C@H](C(C)C)NC(=O)[C@H](C(C)C)N)O JVGDAEKKZKKZFO-RCWTZXSCSA-N 0.000 description 1
- STTYIMSDIYISRG-UHFFFAOYSA-N Valyl-Serine Chemical compound CC(C)C(N)C(=O)NC(CO)C(O)=O STTYIMSDIYISRG-UHFFFAOYSA-N 0.000 description 1
- 241000274427 Xylomelum angustifolium Species 0.000 description 1
- 241000607479 Yersinia pestis Species 0.000 description 1
- 230000009471 action Effects 0.000 description 1
- 239000011543 agarose gel Substances 0.000 description 1
- 108010045023 alanyl-prolyl-tyrosine Proteins 0.000 description 1
- 108010041407 alanylaspartic acid Proteins 0.000 description 1
- 230000029936 alkylation Effects 0.000 description 1
- 238000005804 alkylation reaction Methods 0.000 description 1
- 239000013566 allergen Substances 0.000 description 1
- 108010050025 alpha-glutamyltryptophan Proteins 0.000 description 1
- 229940043376 ammonium acetate Drugs 0.000 description 1
- 235000019257 ammonium acetate Nutrition 0.000 description 1
- 235000011114 ammonium hydroxide Nutrition 0.000 description 1
- 229960000723 ampicillin Drugs 0.000 description 1
- AVKUERGKIZMTKX-NJBDSQKTSA-N ampicillin Chemical compound C1([C@@H](N)C(=O)N[C@H]2[C@H]3SC([C@@H](N3C2=O)C(O)=O)(C)C)=CC=CC=C1 AVKUERGKIZMTKX-NJBDSQKTSA-N 0.000 description 1
- 230000003321 amplification Effects 0.000 description 1
- 238000005349 anion exchange Methods 0.000 description 1
- 125000000129 anionic group Chemical group 0.000 description 1
- 239000003242 anti bacterial agent Substances 0.000 description 1
- 230000000844 anti-bacterial effect Effects 0.000 description 1
- 238000011482 antibacterial activity assay Methods 0.000 description 1
- 229940088710 antibiotic agent Drugs 0.000 description 1
- 230000000890 antigenic effect Effects 0.000 description 1
- 108010007483 arginyl-leucyl-tyrosyl-glutamic acid Proteins 0.000 description 1
- 108010018691 arginyl-threonyl-arginine Proteins 0.000 description 1
- 235000009582 asparagine Nutrition 0.000 description 1
- 229960001230 asparagine Drugs 0.000 description 1
- 235000003704 aspartic acid Nutrition 0.000 description 1
- 108010047857 aspartylglycine Proteins 0.000 description 1
- 108010068265 aspartyltyrosine Proteins 0.000 description 1
- OQFSQFPPLPISGP-UHFFFAOYSA-N beta-carboxyaspartic acid Natural products OC(=O)C(N)C(C(O)=O)C(O)=O OQFSQFPPLPISGP-UHFFFAOYSA-N 0.000 description 1
- 229960002685 biotin Drugs 0.000 description 1
- 235000020958 biotin Nutrition 0.000 description 1
- 239000011616 biotin Substances 0.000 description 1
- 238000006664 bond formation reaction Methods 0.000 description 1
- KGBXLFKZBHKPEV-UHFFFAOYSA-N boric acid Chemical compound OB(O)O KGBXLFKZBHKPEV-UHFFFAOYSA-N 0.000 description 1
- 239000000872 buffer Substances 0.000 description 1
- 210000004900 c-terminal fragment Anatomy 0.000 description 1
- 238000010804 cDNA synthesis Methods 0.000 description 1
- 235000001046 cacaotero Nutrition 0.000 description 1
- AIYUHDOJVYHVIT-UHFFFAOYSA-M caesium chloride Chemical compound [Cl-].[Cs+] AIYUHDOJVYHVIT-UHFFFAOYSA-M 0.000 description 1
- 239000001110 calcium chloride Substances 0.000 description 1
- 229910001628 calcium chloride Inorganic materials 0.000 description 1
- 229940095731 candida albicans Drugs 0.000 description 1
- 125000002091 cationic group Chemical group 0.000 description 1
- 229960004261 cefotaxime Drugs 0.000 description 1
- GPRBEKHLDVQUJE-VINNURBNSA-N cefotaxime Chemical compound N([C@@H]1C(N2C(=C(COC(C)=O)CS[C@@H]21)C(O)=O)=O)C(=O)/C(=N/OC)C1=CSC(N)=N1 GPRBEKHLDVQUJE-VINNURBNSA-N 0.000 description 1
- 238000012512 characterization method Methods 0.000 description 1
- 229920001429 chelating resin Polymers 0.000 description 1
- 239000003795 chemical substances by application Substances 0.000 description 1
- 102000021178 chitin binding proteins Human genes 0.000 description 1
- 108091011157 chitin binding proteins Proteins 0.000 description 1
- 239000012501 chromatography medium Substances 0.000 description 1
- PMMYEEVYMWASQN-IMJSIDKUSA-N cis-4-Hydroxy-L-proline Chemical compound O[C@@H]1CN[C@H](C(O)=O)C1 PMMYEEVYMWASQN-IMJSIDKUSA-N 0.000 description 1
- ARUVKPQLZAKDPS-UHFFFAOYSA-L copper(II) sulfate Chemical compound [Cu+2].[O-][S+2]([O-])([O-])[O-] ARUVKPQLZAKDPS-UHFFFAOYSA-L 0.000 description 1
- 229910000366 copper(II) sulfate Inorganic materials 0.000 description 1
- 238000012258 culturing Methods 0.000 description 1
- ATDGTVJJHBUTRL-UHFFFAOYSA-N cyanogen bromide Chemical compound BrC#N ATDGTVJJHBUTRL-UHFFFAOYSA-N 0.000 description 1
- 230000004665 defense response Effects 0.000 description 1
- 238000004925 denaturation Methods 0.000 description 1
- 230000036425 denaturation Effects 0.000 description 1
- KCFYHBSOLOXZIF-UHFFFAOYSA-N dihydrochrysin Natural products COC1=C(O)C(OC)=CC(C2OC3=CC(O)=CC(O)=C3C(=O)C2)=C1 KCFYHBSOLOXZIF-UHFFFAOYSA-N 0.000 description 1
- 229960003983 diphtheria toxoid Drugs 0.000 description 1
- ZPWVASYFFYYZEW-UHFFFAOYSA-L dipotassium hydrogen phosphate Chemical compound [K+].[K+].OP([O-])([O-])=O ZPWVASYFFYYZEW-UHFFFAOYSA-L 0.000 description 1
- 229910000396 dipotassium phosphate Inorganic materials 0.000 description 1
- BNIILDVGGAEEIG-UHFFFAOYSA-L disodium hydrogen phosphate Chemical compound [Na+].[Na+].OP([O-])([O-])=O BNIILDVGGAEEIG-UHFFFAOYSA-L 0.000 description 1
- 229910000397 disodium phosphate Inorganic materials 0.000 description 1
- VHJLVAABSRFDPM-QWWZWVQMSA-N dithiothreitol Chemical compound SC[C@@H](O)[C@H](O)CS VHJLVAABSRFDPM-QWWZWVQMSA-N 0.000 description 1
- 239000003814 drug Substances 0.000 description 1
- 230000005014 ectopic expression Effects 0.000 description 1
- 238000001962 electrophoresis Methods 0.000 description 1
- 230000000408 embryogenic effect Effects 0.000 description 1
- 238000005516 engineering process Methods 0.000 description 1
- 238000000605 extraction Methods 0.000 description 1
- 230000002349 favourable effect Effects 0.000 description 1
- 239000000835 fiber Substances 0.000 description 1
- 238000001914 filtration Methods 0.000 description 1
- 239000000417 fungicide Substances 0.000 description 1
- 239000007789 gas Substances 0.000 description 1
- 238000010353 genetic engineering Methods 0.000 description 1
- 239000008103 glucose Substances 0.000 description 1
- 235000013922 glutamic acid Nutrition 0.000 description 1
- 239000004220 glutamic acid Substances 0.000 description 1
- 108010085059 glutamyl-arginyl-proline Proteins 0.000 description 1
- 108010073628 glutamyl-valyl-phenylalanine Proteins 0.000 description 1
- VPZXBVLAVMBEQI-UHFFFAOYSA-N glycyl-DL-alpha-alanine Natural products OC(=O)C(C)NC(=O)CN VPZXBVLAVMBEQI-UHFFFAOYSA-N 0.000 description 1
- 108010075431 glycyl-alanyl-phenylalanine Proteins 0.000 description 1
- 108010019407 glycyl-arginyl-glycyl-aspartic acid Proteins 0.000 description 1
- 108010067216 glycyl-glycyl-glycine Proteins 0.000 description 1
- XKUKSGPZAADMRA-UHFFFAOYSA-N glycyl-glycyl-glycine Natural products NCC(=O)NCC(=O)NCC(O)=O XKUKSGPZAADMRA-UHFFFAOYSA-N 0.000 description 1
- 108010033719 glycyl-histidyl-glycine Proteins 0.000 description 1
- 108010050475 glycyl-leucyl-tyrosine Proteins 0.000 description 1
- 108010025801 glycyl-prolyl-arginine Proteins 0.000 description 1
- 108010079413 glycyl-prolyl-glutamic acid Proteins 0.000 description 1
- 108010043293 glycyl-prolyl-glycyl-glycine Proteins 0.000 description 1
- 108010082286 glycyl-seryl-alanine Proteins 0.000 description 1
- PCHJSUWPFVWCPO-UHFFFAOYSA-N gold Chemical compound [Au] PCHJSUWPFVWCPO-UHFFFAOYSA-N 0.000 description 1
- 239000010931 gold Substances 0.000 description 1
- 229910052737 gold Inorganic materials 0.000 description 1
- 239000011544 gradient gel Substances 0.000 description 1
- 239000001963 growth medium Substances 0.000 description 1
- ZJYYHGLJYGJLLN-UHFFFAOYSA-N guanidinium thiocyanate Chemical compound SC#N.NC(N)=N ZJYYHGLJYGJLLN-UHFFFAOYSA-N 0.000 description 1
- 238000003306 harvesting Methods 0.000 description 1
- 239000001307 helium Substances 0.000 description 1
- 229910052734 helium Inorganic materials 0.000 description 1
- SWQJXJOGLNCZEY-UHFFFAOYSA-N helium atom Chemical compound [He] SWQJXJOGLNCZEY-UHFFFAOYSA-N 0.000 description 1
- 108010040030 histidinoalanine Proteins 0.000 description 1
- 244000052637 human pathogen Species 0.000 description 1
- 238000009396 hybridization Methods 0.000 description 1
- 239000001257 hydrogen Substances 0.000 description 1
- 229910052739 hydrogen Inorganic materials 0.000 description 1
- 239000012135 ice-cold extraction buffer Substances 0.000 description 1
- 238000003384 imaging method Methods 0.000 description 1
- 230000028993 immune response Effects 0.000 description 1
- 210000003000 inclusion body Anatomy 0.000 description 1
- 238000011534 incubation Methods 0.000 description 1
- 229960000367 inositol Drugs 0.000 description 1
- CDAISMWEOUEBRE-GPIVLXJGSA-N inositol Chemical compound O[C@H]1[C@H](O)[C@@H](O)[C@H](O)[C@H](O)[C@@H]1O CDAISMWEOUEBRE-GPIVLXJGSA-N 0.000 description 1
- 238000003780 insertion Methods 0.000 description 1
- 230000037431 insertion Effects 0.000 description 1
- 239000002198 insoluble material Substances 0.000 description 1
- 238000001990 intravenous administration Methods 0.000 description 1
- 238000010253 intravenous injection Methods 0.000 description 1
- 230000009545 invasion Effects 0.000 description 1
- BAUYGSIQEAFULO-UHFFFAOYSA-L iron(2+) sulfate (anhydrous) Chemical compound [Fe+2].[O-]S([O-])(=O)=O BAUYGSIQEAFULO-UHFFFAOYSA-L 0.000 description 1
- 229910000359 iron(II) sulfate Inorganic materials 0.000 description 1
- 108010027338 isoleucylcysteine Proteins 0.000 description 1
- 108010078274 isoleucylvaline Proteins 0.000 description 1
- 108010034529 leucyl-lysine Proteins 0.000 description 1
- 108010000761 leucylarginine Proteins 0.000 description 1
- 238000011068 loading method Methods 0.000 description 1
- 125000003588 lysine group Chemical group [H]N([H])C([H])([H])C([H])([H])C([H])([H])C([H])([H])C([H])(N([H])[H])C(*)=O 0.000 description 1
- 108010003700 lysyl aspartic acid Proteins 0.000 description 1
- 108010064235 lysylglycine Proteins 0.000 description 1
- 108010054155 lysyllysine Proteins 0.000 description 1
- 108010017391 lysylvaline Proteins 0.000 description 1
- 229910052943 magnesium sulfate Inorganic materials 0.000 description 1
- 210000004962 mammalian cell Anatomy 0.000 description 1
- SQQMAOCOWKFBNP-UHFFFAOYSA-L manganese(II) sulfate Chemical compound [Mn+2].[O-]S([O-])(=O)=O SQQMAOCOWKFBNP-UHFFFAOYSA-L 0.000 description 1
- 229910000357 manganese(II) sulfate Inorganic materials 0.000 description 1
- 238000004519 manufacturing process Methods 0.000 description 1
- 230000013011 mating Effects 0.000 description 1
- 235000012054 meals Nutrition 0.000 description 1
- 230000007246 mechanism Effects 0.000 description 1
- BCVXHSPFUWZLGQ-UHFFFAOYSA-N mecn acetonitrile Chemical compound CC#N.CC#N BCVXHSPFUWZLGQ-UHFFFAOYSA-N 0.000 description 1
- 229930182817 methionine Natural products 0.000 description 1
- 108010056582 methionylglutamic acid Proteins 0.000 description 1
- 108010068488 methionylphenylalanine Proteins 0.000 description 1
- 238000000520 microinjection Methods 0.000 description 1
- 239000004570 mortar (masonry) Substances 0.000 description 1
- 230000017066 negative regulation of growth Effects 0.000 description 1
- JPXMTWWFLBLUCD-UHFFFAOYSA-N nitro blue tetrazolium(2+) Chemical compound COC1=CC(C=2C=C(OC)C(=CC=2)[N+]=2N(N=C(N=2)C=2C=CC=CC=2)C=2C=CC(=CC=2)[N+]([O-])=O)=CC=C1[N+]1=NC(C=2C=CC=CC=2)=NN1C1=CC=C([N+]([O-])=O)C=C1 JPXMTWWFLBLUCD-UHFFFAOYSA-N 0.000 description 1
- 229920001220 nitrocellulos Polymers 0.000 description 1
- 238000003199 nucleic acid amplification method Methods 0.000 description 1
- 108020004707 nucleic acids Proteins 0.000 description 1
- 102000039446 nucleic acids Human genes 0.000 description 1
- 150000007523 nucleic acids Chemical class 0.000 description 1
- 230000003287 optical effect Effects 0.000 description 1
- 238000004806 packaging method and process Methods 0.000 description 1
- 239000011236 particulate material Substances 0.000 description 1
- 239000013618 particulate matter Substances 0.000 description 1
- 239000008188 pellet Substances 0.000 description 1
- 238000005191 phase separation Methods 0.000 description 1
- 108010082795 phenylalanyl-arginyl-arginine Proteins 0.000 description 1
- 108010070409 phenylalanyl-glycyl-glycine Proteins 0.000 description 1
- 108010065135 phenylalanyl-phenylalanyl-phenylalanine Proteins 0.000 description 1
- 108010024654 phenylalanyl-prolyl-alanine Proteins 0.000 description 1
- 108010024607 phenylalanylalanine Proteins 0.000 description 1
- 108010018625 phenylalanylarginine Proteins 0.000 description 1
- 239000008363 phosphate buffer Substances 0.000 description 1
- 108010025488 pinealon Proteins 0.000 description 1
- 235000021118 plant-derived protein Nutrition 0.000 description 1
- 238000003752 polymerase chain reaction Methods 0.000 description 1
- 239000003910 polypeptide antibiotic agent Substances 0.000 description 1
- 238000011176 pooling Methods 0.000 description 1
- 239000000843 powder Substances 0.000 description 1
- 239000002244 precipitate Substances 0.000 description 1
- 108010087846 prolyl-prolyl-glycine Proteins 0.000 description 1
- 108010079317 prolyl-tyrosine Proteins 0.000 description 1
- 239000012460 protein solution Substances 0.000 description 1
- 230000017854 proteolysis Effects 0.000 description 1
- 230000006337 proteolytic cleavage Effects 0.000 description 1
- ZUFQODAHGAHPFQ-UHFFFAOYSA-N pyridoxine hydrochloride Chemical compound Cl.CC1=NC=C(CO)C(CO)=C1O ZUFQODAHGAHPFQ-UHFFFAOYSA-N 0.000 description 1
- 235000019171 pyridoxine hydrochloride Nutrition 0.000 description 1
- 239000011764 pyridoxine hydrochloride Substances 0.000 description 1
- 230000009257 reactivity Effects 0.000 description 1
- 230000009467 reduction Effects 0.000 description 1
- 238000009877 rendering Methods 0.000 description 1
- 230000010076 replication Effects 0.000 description 1
- 238000011160 research Methods 0.000 description 1
- CDAISMWEOUEBRE-UHFFFAOYSA-N scyllo-inosotol Natural products OC1C(O)C(O)C(O)C(O)C1O CDAISMWEOUEBRE-UHFFFAOYSA-N 0.000 description 1
- 108010048818 seryl-histidine Proteins 0.000 description 1
- 108010069117 seryl-lysyl-aspartic acid Proteins 0.000 description 1
- 108010007375 seryl-seryl-seryl-arginine Proteins 0.000 description 1
- 108010071207 serylmethionine Proteins 0.000 description 1
- 239000002002 slurry Substances 0.000 description 1
- AJPJDKMHJJGVTQ-UHFFFAOYSA-M sodium dihydrogen phosphate Chemical compound [Na+].OP(O)([O-])=O AJPJDKMHJJGVTQ-UHFFFAOYSA-M 0.000 description 1
- 239000011684 sodium molybdate Substances 0.000 description 1
- TVXXNOYZHKPKGW-UHFFFAOYSA-N sodium molybdate (anhydrous) Chemical compound [Na+].[Na+].[O-][Mo]([O-])(=O)=O TVXXNOYZHKPKGW-UHFFFAOYSA-N 0.000 description 1
- 239000001488 sodium phosphate Substances 0.000 description 1
- 238000005063 solubilization Methods 0.000 description 1
- 230000007928 solubilization Effects 0.000 description 1
- 230000003019 stabilising effect Effects 0.000 description 1
- 239000003381 stabilizer Substances 0.000 description 1
- 230000000087 stabilizing effect Effects 0.000 description 1
- 238000010186 staining Methods 0.000 description 1
- 239000012192 staining solution Substances 0.000 description 1
- 238000010561 standard procedure Methods 0.000 description 1
- 108010058363 sterol carrier proteins Proteins 0.000 description 1
- 239000000126 substance Substances 0.000 description 1
- 239000000758 substrate Substances 0.000 description 1
- 239000008362 succinate buffer Substances 0.000 description 1
- 239000000725 suspension Substances 0.000 description 1
- 238000003786 synthesis reaction Methods 0.000 description 1
- 230000001225 therapeutic effect Effects 0.000 description 1
- 229960003495 thiamine Drugs 0.000 description 1
- DPJRMOMPQZCRJU-UHFFFAOYSA-M thiamine hydrochloride Chemical compound Cl.[Cl-].CC1=C(CCO)SC=[N+]1CC1=CN=C(C)N=C1N DPJRMOMPQZCRJU-UHFFFAOYSA-M 0.000 description 1
- 235000019190 thiamine hydrochloride Nutrition 0.000 description 1
- 239000011747 thiamine hydrochloride Substances 0.000 description 1
- 230000002588 toxic effect Effects 0.000 description 1
- 231100000419 toxicity Toxicity 0.000 description 1
- 230000001988 toxicity Effects 0.000 description 1
- 230000005030 transcription termination Effects 0.000 description 1
- 230000002103 transcriptional effect Effects 0.000 description 1
- 238000011426 transformation method Methods 0.000 description 1
- LENZDBCJOHFCAS-UHFFFAOYSA-N tris Chemical compound OCC(N)(CO)CO LENZDBCJOHFCAS-UHFFFAOYSA-N 0.000 description 1
- RYFMWSXOAZQYPI-UHFFFAOYSA-K trisodium phosphate Chemical compound [Na+].[Na+].[Na+].[O-]P([O-])([O-])=O RYFMWSXOAZQYPI-UHFFFAOYSA-K 0.000 description 1
- 108010038745 tryptophylglycine Proteins 0.000 description 1
- 108010035534 tyrosyl-leucyl-alanine Proteins 0.000 description 1
- 108010051110 tyrosyl-lysine Proteins 0.000 description 1
- 210000003934 vacuole Anatomy 0.000 description 1
- 108010015385 valyl-prolyl-proline Proteins 0.000 description 1
- 108010009962 valyltyrosine Proteins 0.000 description 1
- 229940011671 vitamin b6 Drugs 0.000 description 1
- 238000002424 x-ray crystallography Methods 0.000 description 1
- NWONKYPBYAMBJT-UHFFFAOYSA-L zinc sulfate Chemical compound [Zn+2].[O-]S([O-])(=O)=O NWONKYPBYAMBJT-UHFFFAOYSA-L 0.000 description 1
- 229910000368 zinc sulfate Inorganic materials 0.000 description 1
- 239000011686 zinc sulphate Substances 0.000 description 1
Images
Classifications
-
- C—CHEMISTRY; METALLURGY
- C12—BIOCHEMISTRY; BEER; SPIRITS; WINE; VINEGAR; MICROBIOLOGY; ENZYMOLOGY; MUTATION OR GENETIC ENGINEERING
- C12N—MICROORGANISMS OR ENZYMES; COMPOSITIONS THEREOF; PROPAGATING, PRESERVING, OR MAINTAINING MICROORGANISMS; MUTATION OR GENETIC ENGINEERING; CULTURE MEDIA
- C12N15/00—Mutation or genetic engineering; DNA or RNA concerning genetic engineering, vectors, e.g. plasmids, or their isolation, preparation or purification; Use of hosts therefor
- C12N15/09—Recombinant DNA-technology
- C12N15/11—DNA or RNA fragments; Modified forms thereof; Non-coding nucleic acids having a biological activity
-
- C—CHEMISTRY; METALLURGY
- C12—BIOCHEMISTRY; BEER; SPIRITS; WINE; VINEGAR; MICROBIOLOGY; ENZYMOLOGY; MUTATION OR GENETIC ENGINEERING
- C12N—MICROORGANISMS OR ENZYMES; COMPOSITIONS THEREOF; PROPAGATING, PRESERVING, OR MAINTAINING MICROORGANISMS; MUTATION OR GENETIC ENGINEERING; CULTURE MEDIA
- C12N15/00—Mutation or genetic engineering; DNA or RNA concerning genetic engineering, vectors, e.g. plasmids, or their isolation, preparation or purification; Use of hosts therefor
- C12N15/09—Recombinant DNA-technology
- C12N15/63—Introduction of foreign genetic material using vectors; Vectors; Use of hosts therefor; Regulation of expression
- C12N15/79—Vectors or expression systems specially adapted for eukaryotic hosts
- C12N15/82—Vectors or expression systems specially adapted for eukaryotic hosts for plant cells, e.g. plant artificial chromosomes (PACs)
- C12N15/8241—Phenotypically and genetically modified plants via recombinant DNA technology
- C12N15/8261—Phenotypically and genetically modified plants via recombinant DNA technology with agronomic (input) traits, e.g. crop yield
- C12N15/8271—Phenotypically and genetically modified plants via recombinant DNA technology with agronomic (input) traits, e.g. crop yield for stress resistance, e.g. heavy metal resistance
- C12N15/8279—Phenotypically and genetically modified plants via recombinant DNA technology with agronomic (input) traits, e.g. crop yield for stress resistance, e.g. heavy metal resistance for biotic stress resistance, pathogen resistance, disease resistance
- C12N15/8282—Phenotypically and genetically modified plants via recombinant DNA technology with agronomic (input) traits, e.g. crop yield for stress resistance, e.g. heavy metal resistance for biotic stress resistance, pathogen resistance, disease resistance for fungal resistance
-
- A—HUMAN NECESSITIES
- A01—AGRICULTURE; FORESTRY; ANIMAL HUSBANDRY; HUNTING; TRAPPING; FISHING
- A01N—PRESERVATION OF BODIES OF HUMANS OR ANIMALS OR PLANTS OR PARTS THEREOF; BIOCIDES, e.g. AS DISINFECTANTS, AS PESTICIDES OR AS HERBICIDES; PEST REPELLANTS OR ATTRACTANTS; PLANT GROWTH REGULATORS
- A01N65/00—Biocides, pest repellants or attractants, or plant growth regulators containing material from algae, lichens, bryophyta, multi-cellular fungi or plants, or extracts thereof
- A01N65/08—Magnoliopsida [dicotyledons]
-
- A—HUMAN NECESSITIES
- A01—AGRICULTURE; FORESTRY; ANIMAL HUSBANDRY; HUNTING; TRAPPING; FISHING
- A01N—PRESERVATION OF BODIES OF HUMANS OR ANIMALS OR PLANTS OR PARTS THEREOF; BIOCIDES, e.g. AS DISINFECTANTS, AS PESTICIDES OR AS HERBICIDES; PEST REPELLANTS OR ATTRACTANTS; PLANT GROWTH REGULATORS
- A01N65/00—Biocides, pest repellants or attractants, or plant growth regulators containing material from algae, lichens, bryophyta, multi-cellular fungi or plants, or extracts thereof
- A01N65/08—Magnoliopsida [dicotyledons]
- A01N65/20—Fabaceae or Leguminosae [Pea or Legume family], e.g. pea, lentil, soybean, clover, acacia, honey locust, derris or millettia
-
- A—HUMAN NECESSITIES
- A01—AGRICULTURE; FORESTRY; ANIMAL HUSBANDRY; HUNTING; TRAPPING; FISHING
- A01N—PRESERVATION OF BODIES OF HUMANS OR ANIMALS OR PLANTS OR PARTS THEREOF; BIOCIDES, e.g. AS DISINFECTANTS, AS PESTICIDES OR AS HERBICIDES; PEST REPELLANTS OR ATTRACTANTS; PLANT GROWTH REGULATORS
- A01N65/00—Biocides, pest repellants or attractants, or plant growth regulators containing material from algae, lichens, bryophyta, multi-cellular fungi or plants, or extracts thereof
- A01N65/40—Liliopsida [monocotyledons]
- A01N65/44—Poaceae or Gramineae [Grass family], e.g. bamboo, lemon grass or citronella grass
-
- A—HUMAN NECESSITIES
- A61—MEDICAL OR VETERINARY SCIENCE; HYGIENE
- A61P—SPECIFIC THERAPEUTIC ACTIVITY OF CHEMICAL COMPOUNDS OR MEDICINAL PREPARATIONS
- A61P31/00—Antiinfectives, i.e. antibiotics, antiseptics, chemotherapeutics
- A61P31/04—Antibacterial agents
-
- A—HUMAN NECESSITIES
- A61—MEDICAL OR VETERINARY SCIENCE; HYGIENE
- A61P—SPECIFIC THERAPEUTIC ACTIVITY OF CHEMICAL COMPOUNDS OR MEDICINAL PREPARATIONS
- A61P31/00—Antiinfectives, i.e. antibiotics, antiseptics, chemotherapeutics
- A61P31/10—Antimycotics
-
- C—CHEMISTRY; METALLURGY
- C07—ORGANIC CHEMISTRY
- C07K—PEPTIDES
- C07K14/00—Peptides having more than 20 amino acids; Gastrins; Somatostatins; Melanotropins; Derivatives thereof
- C07K14/415—Peptides having more than 20 amino acids; Gastrins; Somatostatins; Melanotropins; Derivatives thereof from plants
-
- A—HUMAN NECESSITIES
- A61—MEDICAL OR VETERINARY SCIENCE; HYGIENE
- A61K—PREPARATIONS FOR MEDICAL, DENTAL OR TOILETRY PURPOSES
- A61K38/00—Medicinal preparations containing peptides
Definitions
- This invention relates to isolated proteins which exert inhibitory activity on the growth of fungi and bacteria, which fungi and bacteria include some microbial pathogens of plants and animals.
- the invention also relates to recombinant genes which include sequences encoding the proteins, the expression products of which recombinant genes can contribute to plant cells or cells of other organism's defence against invasion by microbial pathogens.
- the invention further relates to the use of the proteins and/or genes encoding the proteins for the control of microbes in human and veterinary clinical conditions.
- Microbial diseases of plants are a significant problem to the agricultural and horticultural industries. Plant diseases in general cause millions of tonnes of crop losses annually with fungal and bacterial diseases responsible for significant portions of these losses.
- One possible way of combating fungal and bacterial diseases is to provide transgenic plants capable of expressing a protein or proteins which in some way increase the resistance of the plant to pathogen attack.
- a simple strategy is to first identify a protein with antimicrobial activity in vitro, to clone or synthesise the DNA sequence encoding the protein, to make a chimaeric gene construct for efficient expression of the protein in plants, to transfer this gene to transgenic plants and to assess the effect of the introduced gene on resistance to microbial pathogens by comparison with control plants.
- antimicrobial proteins for engineering disease resistance in transgenic plants
- highly potent antimicrobial proteins can be used for the control of plant disease by direct application (De Bolle, M. F. C. et al. [1993] in Mechanisms of Plant Defense Responses, B. Fritig and M. Legrand eds., Kluwer Acad. Publ., Dordrecht, NL, pp. 433-436).
- antimicrobial peptides have potential therapeutic applications in human and veterinary medicine. Although this has not been described for peptides of plant origin it is being actively explored with peptides from animals and has reached clinical trials (Jacob, L. and Zasloff, M. [1994] in “Antimicrobial Peptides”, CIBA Foundation Symposium 186, John Wiley and Sons Publ., Chichester, UK, pp. 197-223).
- Antimicrobial proteins exhibit a variety of three-dimensional structures which will determine in large part the activity which they manifest. Many of the global structures exhibited by these proteins have been determined (Broekaert W.F. et a. (1997) Crit. Rev. in Plant Sci. 16(3):297-323). A large factor in determining the stability of these proteins is the presence of disulfide bridges between various cysteines located in a helical and ⁇ -sheet regions. Many peptides with toxic activity such as conotoxin are well known to be stabilized by disulfide bridges (see for example Hill, J. M. et al. (1996) Biochemistry 35(27): 8824-8835).
- a compact structure is formed consisting of a helix, a small -hairpin, a cis-hydroxyproline, and several turns.
- the molecule is stabilized by three disulfide bonds, two of which connect the ⁇ -helix and the ⁇ -sheet, forming a solid structural core.
- eight arginine and lysine side chains in this molecule project into the solvent in a radial orientation relative to the core of the molecule.
- These cationic side chains form potential sites of interaction with anionic sites on pathogen membranes (Hill, J. M. et al. supra).
- Macadamia integrifolia (Mi) seeds or from cotton or cocoa seeds.
- protein fragments which are antifungal can be derived from larger seed storage proteins containing regions of substantial similarity to the antimicrobial proteins from macadamia described here. Examples of seed storage proteins which contain regions similar to the proteins which have been purified can be seen in FIG. 4.
- Macadamia integrifolia belongs to the family Proteaceae. M. integrifolia, also known as Bauple Nut or Queensland Nut, is considered by some to be the world's best edible nut.
- Cotton Gossypium hirsutum ) belongs to the family Malvaceae and is cultivated extensively for its fiber. Cocoa ( Threobroma cacao ) belongs to the family Sterculiaceae and is used around the world for a wide variety of cocoa products.
- a protein fragment having antimicrobial activity wherein said protein fragment is selected from:
- a protein containing at least one polypeptide fragment according to the first embodiment wherein said polypeptide fragment has a sequence selected from within a sequence comprising SEQ ID NO: 1, SEQ ID NO: 3 or SEQ ID NO: 5.
- a fourth embodiment of the invention there is provided an isolated or synthetic DNA encoding a protein according to the first embodiment
- a DNA construct which includes a DNA according to the fourth embodiment operatively linked to elements for the expression of said encoded protein.
- transgenic plant harbouring a DNA construct according to the fifth embodiment.
- reproductive material of a transgenic plant according to the sixth embodiment is provided.
- composition comprising an antimicrobial protein according to the first embodiment together with an agriculturally-acceptable carrier diluent or excipient.
- composition comprising an antimicrobial protein according to the first embodiment together with an pharmaceutically-acceptable carrier diluent or excipient.
- a method of controlling microbial infestation of a plant comprising:
- an eleventh embodiment of the invention there is provided a method of controlling microbial infestation of a mammalian animal, the method comprising treating the animal with an antimicrobial protein according to the first embodiment or a composition according to the ninth embodiment.
- a method of preparing an antimicrobial protein which method comprises the steps of:
- inventions include methods for producing antimicrobial protein.
- FIG. 1 shows the results of cation-exchange chromatography of the basic protein fraction of a Macadamia integrifolia extract with the results of a bioassay for antimicrobial activity shown for fractions in the region of MiAMP2c elution.
- FIG. 2 shows the results of including 1 mM Ca 2+ in a parallel bioassay of fractions from the cation-exchange separation.
- FIG. 3 shows a reverse-phase HPLC profile of highly inhibitory fractions containing MiAMP2c from the cation-exchange separation in FIGS. 1 and 2 together with % growth inhibition exhibited by the HPLC fractions.
- FIG. 4 shows the amino acid sequences of MiAMP2a, b, c and d and protein fragments derived from other seed storage proteins which contain regions of homology to the MiAMP2 series of antimicrobial proteins.
- FIG. 5 shows an example of a synthetic nucleotide sequence which can be used for the expression and secretion of MiAMP2c in transgenic plants.
- FIG. 6 shows the alignment of clones 1-3 from macadamia containing MiAMP2a, b, c and d subunits together with sequences from cocoa and cotton vicilin seed storage proteins which exhibit significant homology to the macadamia clones.
- FIG. 7 displays a series of secondary structure predictions for MiAMP2c.
- FIG. 8 shows a three-dimensional model of the MiAMP2c protein.
- FIG. 9 shows stained SDS-PAGE gels of protein fractions at various stages in the expression and purification of TcAMP1 ( Theobroma cacao subunit 1), MiAMP2a, MiAMP2b, MiAMP2c and MiAMP2d expressed in E. coli liquid culture.
- FIG. 10 shows the reverse-phase HPLC purification of cocoa subunit 2 (TcAMP2) after the initial purification step using Ni-NTA media.
- FIG. 11 shows a western blot of crude protein extracts from various plant species using rabbit antiserum raised to MiAMP2c.
- FIG. 12 shows a cation-exchange fractionation of the Stenocarpus sinuatus basic protein fraction along with the accompanying western blot which shows the presence of immunologically-related proteins in a range of fractions.
- FIG. 13 shows a reverse-phase HPLC separation of Stenocarpus sinuatus cation-exchange fractions which had previously reacted with MiAMP2c antibodies (see FIG. 14).
- a western blot is also presented which reveals the presence of putative MiAMP2c homologues in individual HPLC fractions.
- FIG. 14 is a map of the binary vector pPCV91-MiAMP2c as an example of a vector that can be used to express these antimicrobial proteins in transgenic plants.
- FIG. 15 shows a western blot to detect MiAMP2c expressed in transgenic tobacco plants.
- homologue is used herein to denote any polypeptide having substantial similarity in composition and sequence to the polypeptide used as the reference.
- the homologue of a reference polypeptide will contain key elements such as cysteine or other residues spaced at identical intervals such that a substantially similar three-dimensional global structure is adopted by the homologue as compared to the reference.
- the homologue will also exhibit substantially the same antimicrobial activity as the reference protein.
- the present inventors have identified a new class of proteins with antimicrobial activity.
- Prototype proteins can be isolated from seeds of Macadamia integrifolia.
- the invention thus provides antimicrobial proteins per se and also DNA sequences encoding these antimicrobial proteins.
- the invention also provides amino acid sequences of proteins which are homologous to the prototype antimicrobial proteins from Macadamia integrifolia.
- this invention also provides amino acid sequences of homologues from other species which have hitherto been unrecognized as having antimicrobial activity.
- MiAMP2a fragments contained in the 666-amino-acid clone are termed MiAMP2a, b and d as per their locations in the cloned nucleotide sequence.
- Several other sequences with significant homology to the MiAMP2a, b, c, and d protein fragments were then identifed in the Entrez data base. These homologous sequences were contained within larger seed storage proteins from cotton and cocoa which sequences had not been previously described as containing antimicrobial protein sequences or as exhibiting antimicrobial activity. Fragments of larger seed storage proteins containing sequences homologous to MiAMP2c were tested and are here demonstrated to exhibit antimicrobial activity. Thus, the inventors have established a process for obtaining antimicrobial protein fragments from larger seed storage proteins. In the light of these findings, it is evident that fragments of other seed storage proteins containing sequences similar to the proteins described will also exhibit antimicrobial activity.
- the 47-amino-acid TcAMP1 for Theobroma cacao antimicrobial protein 1
- the 60-amino-acid TcAMP2 sequences were derived from a cocoa vicilin seed storage protein gene sequence (which contains 525 amino acids) (Spencer, M. E. and Hodge R. [1992] Planta 186:567-576). These derived fragments were then expressed in liquid culture. Cocoa vicilin fragments thus expressed and subsequently purified (Examples 10 and 11), were shown to be antimicrobial (Example 15). This is the first report that fragments of the cocoa vicilin protein possess antimicrobial activity.
- cocoa-vicilin-derived fragments exhibit antimicrobial activity
- additional proteins which exhibit antimicrobial activity.
- proteins from Stenocarpus sinuatus which are of similar size to MiAMP2 subunits, react with MiAMP2c antiserum, and contain sequences homologous to MiAMP2 proteins (see FIG. 4).
- sequences homologous to the MiAMP2c subunit i.e., MiAMP2a, b, d; TcAMP1; TcAMP2; and cotton fragments 1, 2 and 3—see FIG. 4 constitute proteins which contain the fragment with antimicrobial activity.
- the proteins which contain regions of sequence homologous to MiAMP2 can be used to construct nucleotide sequences encoding 1) the active fragments of larger proteins, or 2) fusions of multiple antimicrobial fragments. This can be done using standard codon tables and cloning methods as described in laboratory manuals such as Current Protocols in Molecular Biology (copyright 1987-1995 edited by Ausubel F. M. et al. and published by John Wiley & Sons, Inc., printed in the USA). Subsequently, these can be expressed in liquid culture for purification and testing, or the sequences can be expressed in transgenic plants after placing them in appropriate expression vectors.
- the antimicrobial proteins per se will manifest a particular three-dimensional structure which may be determined using X-ray crystallography or nuclear magnetic resonance techniques. This structure will be responsible in large part for the antimicrobial activity of the protein.
- the sequence of the protein can also be subjected to structure prediction algorithms to assess whether any secondary structure elements are likely to be exhibited by the protein (see Example 8 and FIG. 7). Secondary structures, thus predicted, can then be used to model three-dimensional global structures. Although three-dimensional structure prediction is not feasible for most proteins, the secondary structure predictions for MiAMP2c were sufficiently simple and clear that a three-dimensional model structure has been obtained for the MiAMP2c protein. Homologues exhibiting the same cysteine spacing and other key elements will also adopt the same three-dimensional structure.
- Example 8 shows that the structure most likely to be adopted by MiAMP2c (and homologues) is a helix-turn-helix structure stabilised by at least two disulfide bridges connecting the two antiparallel ⁇ -helical segments (see FIG. 8). Additional stabilisation can be provided by an extra disulfide bridge (e.g., as in MiAMP2b) or by a hydrophobic ring-stacking interaction between tyrosine and/or phenylalanine residues (e.g., MiAMP2a and MiAMP2c), each located on the same face of the ⁇ -helical segments as the normally present cysteine residues which participate in the 2 disulfide linkages mentioned above. NMR signals exhibited by MiAMP2c are consistent with the three-dimensional global model produced from the secondary-structure predictions mentioned above.
- cysteine residues reside on one face of the helix in which they are contained.
- Aromatic tyrosine (or phenylalanine) residues can also function to add stability to the protein structure if they are located on the same face of the helix as the cysteine side chains. This can be accomplished by providing appropriate spacing of two or three residues between the aromatic residue and the proximate cysteine residue (i.e., Z-X-X-C-X-X-X-C-nX-C-X-X-X-C-X-X-Z where Z is tyrosine or phenylalanine).
- the distribution of positive (and negative) charges on the various surfaces of the protein will also serve a critical role in determining the structure and activity of the protein.
- the distribution of positively-charged residues in an ⁇ -helical region of a protein can result in positive charges lying on one face of the helix or may result in the charged residues being concentrated in some particular portion of the molecule.
- An alternative distribution of positively charged residues is for them to project into the solvent in a radial orientation to the core of the protein. This orientation is predicted for several of the MiAMP2 homologues (data not shown).
- the spacing which is required for positioning of the residues on one face of the helix or the-spacing required to accomplish a radial orientation from the core can easily be determined by one skilled in the art using a helical wheel plot with the sequence of interest.
- a helical wheel plot uses the fact that, in ⁇ -helices, each turn of the helix is composed of 3.6 residues on average. This number translates to 100° of rotational translation per residue making it possible to construct a plot showing the distribution of side chains in a helical region.
- FIG. 8 shows how the spacing of charged residues can lead to most of the positively charged side chains being localised on one face of the helix. It will be appreciated by one of skill in the art that positive charges are conferred by arginine and lysine residues.
- DNA sequences reported here are an extremely powerful tool which can be used to obtain homologous genes from other species.
- DNA sequences one skilled in the art can design and synthesise oligonucleotide probes which can be used to screen cDNA libraries from other species of plants for the presence of genes encoding antimicrobial proteins homologous to the ones described here. This would simply involve construction of a cDNA library and subsequent screening of the library using as the oligonucleotide probe one or part of one of the sequences reported here (such as sequence ID. No. 2 or the PCR fragment described in Example 9).
- oligonucleotide sequences coding for proteins homologous to MiAMP2 can also be used for this purpose (e.g., DNA sequences corresponding to cotton and cocoa vicilins).
- Making and screening of a cDNA library can be carried out by purchasing a kit for said purpose (e.g., from Stratagene) or by following well established protocols described in available DNA cloning manuals (see Current Protocols in Molecular Biology, supra). It is relatively straight forward to construct libraries of various species and to specifically isolate vicilin homologues which are similar to the Macadamia, cotton, or cocoa vicilins by using a simple DNA hybridization technique to screen such libraries.
- these vicilin-related sequences can then be examined for the presence of MiAMP2-like subunits.
- Such subunits can easily be expressed in E. coli using the system described in Examples 10 and 11. Subsequently, these proteins can also be expressed in transgenic.
- Genes, or fragments thereof, under the control of a constitutive or inducible promoter, can then be cloned into a biological system which allows expression of the protein encoded thereby. Transformation methods allowing for the protein to be expressed in a variety of systems are known. The protein can thus be expressed in any suitable system for the purpose of producing the protein for further use. Suitable hosts for the expression of the protein include E. coli, fungal cells, insect cells, mammalian cells, and plants. Standard methods for expressing proteins in such hosts are described in a variety of texts including section 16 (Protein Expression) of Current Protocols in Molecular Biology (supra).
- Plant cells can be transformed with DNA constructs of the invention according to a variety of known methods (Agrobacterium, Ti plasmids, electroporation, micro-injections, micro-projectile gun, and the like).
- DNA sequences encoding the Macadamia integrifolia antimicrobial protein subunits (i.e. fragments a, b, c, or d from the MiAMP2 clones) as well as DNA coding for other homologues can be used in conjunction with a DNA sequence encoding a preprotein from which the mature protein is produced.
- This preprotein can contain a native or synthetic signal peptide sequence which will target the protein to a particular cell compartment (e.g., the apoplast or the vacuole).
- These coding sequences can be ligated to a plant promoter sequence that will ensure strong expression in plant cells.
- This promoter sequence might ensure strong constitutive expression of the protein in most or all plant cells, it may be a promoter which ensures expression in specific tissues or cells that are susceptible to microbial infection and it may also be a promoter which ensures strong induction of expression during the infection process.
- These types of gene cassettes will also include a transcription termination and polyadenylation sequence 3′ of the antimicrobial protein coding region to ensure efficient production and stabilisation of the mRNA encoding the antimicrobial proteins. It is possible that efficient expression of the antimicrobial proteins disclosed herein might be facilitated by inclusion of their individual DNA sequences into a sequence encoding a much larger protein which is processed in planta to produce one or more active MiAMP2-like fragments.
- Gene cassettes encoding the MiAMP2 series antimicrobial proteins i.e., MiAMP2a, b, c, or d; or all of the subunits together; or the entire MiAMP2 clone
- the gene cassettes can then be expressed in plant cells using two common methods. Firstly, the gene cassettes can be ligated into binary vectors carrying: i) left and right border sequences that flank the T-DNA of the Agrobacterium tumefaciens Ti plasmid; ii) a suitable selectable marker gene for the selection of antibiotic resistant plant cells; iii) origins of replication that function in either A.
- This binary vector carrying the chimaeric MiAMP2 encoding gene can be introduced by either electroporation or triparental mating into A. tumefaciens strains carrying disarmed Ti plasmids such as strains LBA4404, GV3101, and AGL1 or into A. rhizogenes strains such as A4 or NCCP1885. These Agrobacterium strains can then be co-cultivated with suitable plant explants or intact plant tissue and the transformed plant cells and/or regenerants selected using antibiotic resistance.
- a second method of gene transfer to plants can be achieved by direct insertion of the gene in target plant cells.
- an MiAMP2-encoding gene cassette can be co-precipitated onto gold or tungsten particles along with a plasmid encoding a chimaeric gene for antibiotic resistance in plants.
- the tungsten particles can be accelerated using a fast flow of helium gas and the particles allowed to bombard a suitable plant tissue.
- This can be an embryogenic cell culture, a plant explant, a callus tissue or cell suspension or an intact meristem. Plants can be recovered using the antibiotic resistance gene for selection and antibodies used to detect plant cells expressing the MiAMP2 proteins or related fragments.
- MiAMP2 proteins in the transgenic plants can be detected using either antibodies raised to the protein(s) or using antimicrobial bioassays. These and other related methods for the expression of MiAMP2 proteins or fragments thereof in plants are described in Plant Molecular Biology (2nd ed., edited by Gelvin, S. B. and Schilperoort, R. A., ⁇ 1994, published by Kluwer Academic Publishers, Dordrecht, The Netherlands)
- Both monocotyledonous and dicotyledonous plants can be transformed and regenerated.
- genetically modified plants include maize, banana, peanut, field peas, sunflower, tomato, canola, tobacco, wheat, barley, oats, potato, soybeans, cotton, carnations, roses, sorghum.
- These, as well as other agricultural plants can be transformed with the antimicrobial genes such that they would exhibit a greater degree of resistance to pathogen attack.
- the proteins can be used for the control of diseases by topological application.
- the invention also relates to application of antimicrobial protein in the control of pathogens of mammals, including humans.
- the protein can be used either in topological or intravenous applications for the control of microbial infections.
- the invention includes within its scope the preparation of antimicrobial proteins based on the prototype MiAMP2 series of proteins.
- New sequences can be designed from the MiAMP2 amino acid sequences which substantially retain the distribution of positively charged residues relative to cysteine residues as found in the MiAMP2 proteins.
- the new sequence can be synthesised or expressed from a gene encoding the sequence in an appropriate host cell. Suitable methods for such procedures have been described above. Expression of the new protein in a genetically engineered cell will typically result in a product having a correct three-dimensional structure, including correctly formed disulphide linkages between cysteine residues.
- MiAMP2 series of antimicrobial proteins or MiAMP2 proteins
- Each protein fragment of the series has a characteristic pI value.
- MiAMP2a, b, c, and d subunits as shown in FIG.
- the proteins have predicted pI values of 4.4, 4.6, 11.5, and 11.6 respectively (predicted using raw sequence data without the His tag or cleavage sequences associated with expression of fragments in the vector pET16b), and contain two sets of CXXXC motifs which are important in stabilising the three-dimensional structure of the protein through the formation of disulfide bonds. Additionally, the proteins contain either an added set of aromatic (tyrosine/phenylalanine) residues or an added set of cysteine residues located at positions which would give more stability to the helix-turn-helix structure as described above and in Example 8.
- MiAMP2a, b, c and d sequences exhibit significant similarity with regions of cocoa vicilin and cotton vicilin (as seen in FIG. 6). Some similarity is also seen with fragments from other seed storage proteins of peanut (Burks, A. W. et al. [1995] J. Clin. Invest. 96 (4), 1715-1721), maize (Belanger, F.
- both cotton and cocoa vicilin-derived subunits retain the conserved tyrosine or phenylalanine residues as additional stabilisers of the tertiary structure.
- the cotton and cocoa vicilins with 525 and 590 amino acids, respectively, are much larger proteins than MiAMP2c (47 amino acids) (see FIGS. 4 and 6).
- MiAMP2 subunits also share some homology with MBP-1 antimicrobial protein from maize (Duvick, J. P. et al. (1992) J Biol Chem 267:18814-20) the number of residues between the CXXXC motifs is 13 which puts MBP-1 outside the specifications for the spacing given here in this application.
- MBP-1 is also a smaller protein (33 amino acids), overall, than the sequences claimed here and there is no evidence available the MBP-1 is derived from a larger seed storage protein other than some similarity with a portion of miaze globulin protein.
- MBP-1 cannot be derived from from the maize globulin since maize globulin contains 10 residues between the two CXXXC motifs while MBP-1 contains 13.
- the alignments in FIGS. 4 and 6 show the similarity in cysteine spacing between MiAMP2 subunits and the cocoa and cotton vicilin-derived molecules. The cysteine and the aromatic tyrosine/phenylalanine residues in FIGS. 4 and 6 are highlighted with bold underlined text.
- FIG. 4 also shows the alignment of additional proteins which can be expressed in liquid culture and shown to exhibit antimicrobial activity.
- MiAMP2 homologues exhibit antifungal activity.
- MiAMP2 homologues show very significant inhibition of fungal growth at concentrations as low as 2 ⁇ g/ml for some of the pathogens/microbes against which the proteins were tested. Thus they can be used to provide protection against several plant diseases.
- MiAMP2 homologues can be used as fungicides or antibiotics by application to plant parts.
- the proteins can also be used to inhibit growth of pathogens by expressing them in transgenic plants.
- the proteins can also be used for the control of human pathogens by topological application or intravenous injection.
- One characteristic of the proteins is that inhibition of some microbes is suppressed by the presence of Ca 2+ (1 mM). An example of this effect is provided for MiAMP2c subunit in Table 1.
- MiAMP2 proteins and homologues could also function as insect control agents. Since some of the proteins are extremely basic (e.g., pI>11.5 for MiAMP2c and d subunits), they would maintain a strong net-positive charge even in the highly alkaline environment of an insect gut. This strong net-positive charge would enable it to interact with negatively charged structures within the gut. This interaction may lead to inefficient feeding, slowing of growth, and possibly death of the insect pest.
- the resulting homogenate was run through a kitchen strainer to remove larger particulate material and then further clarified by centrifugation (4000 rpm for 15 min) in a large capacity centrifuge. Solid ammonium sulphate was added to the supernatant to obtain 30% relative saturation and the precipitate allowed to form overnight with stirring at 4° C. Following centrifugation at 4000 rpm for 30 min, the supernatant was taken and ammonium sulphate added to achieve 70% relative saturation. The solution was allowed to precipitate overnight and then centrifuged at 4000 rpm for 30 min in order to collect the precipitated protein fraction.
- the precipitated protein was resuspended in a minimal volume of extraction buffer and centrifuged once again (13,000 rpm ⁇ 30 min) to remove the any insoluble material yet remaining.
- dialysis (10 mM ethanolamine pH 9.0, 2 mM EDTA and 1 mM PMSF) to remove residual ammonium sulphate
- the protein solution was passed through a Q-Sepharose Fast Plow column (5 ⁇ 12 cm) previously equilibrated with 10 mM ethanolamine (pH 9), 2 mM in EDTA).
- the collected flowthrough from this column represents the basic (pI>9) protein fraction of the seeds. This fraction was further purified as described in Example 3.
- bioassays to assess antifungal and antibacterial activity were carried out in 96-well microtitre plates.
- the test organism was suspended in a synthetic growth medium consisting of K 2 HPO 4 (2.5 mM), MgSO 4 (50 ⁇ M), CaCl 2 (50 ⁇ M), FeSO 4 (5 ⁇ M), CoCl 2 (0.1 ⁇ M), CuSO 4 (0.1 ⁇ M), Na 2 MoO 4 (2 ⁇ M), H 3 BO 3 (0.5 ⁇ M), KI (0.1 ⁇ M), ZnSO 4 (0.5 ⁇ M), MnSO 4 (0.1 ⁇ M), glucose (10 g/L), asparagine (1 g/L), methionine (20 mg/L), myo-inositol (2 mg/L), biotin (0.2 mg/L), thiamine-HCl (1 mg/L) and pyridoxine-HCl (0.2 mg/L).
- test organism consisted of bacterial cells, fungal spores (50,000 spores/ml) or fungal mycelial fragments (produced by blending a hyphal mass from a culture of the fungus to be tested and then filtering through a fine mesh to remove larger hyphal masses). Fifty microlitres of the test organism suspended in medium was placed into each well of the microtitre plate. A further 50 ⁇ l of the test antimicrobial solution was added to appropriate wells. To deal with well-to-well variability in the bioassay, 4 replicates of each test solution were done. Sixteen wells from each 96-well plate were used as controls for comparison with the test solutions.
- incubation was at 25° C. for 48 hours. All fungi including yeast were grown at 25° C. E. coli were grown at 37° C. and other bacteria were bioassayed at 28° C. Percent growth inhibition was measured by following the absorbance at 600 nm of growing cultures over various time intervals and is defined as 100 times the ratio of the average change in absorbance in the control wells minus the change in absorbance in the test well divided by the average change in absorbance at 600 nm for the control wells (i.e., [(avg change in control wells ⁇ change in test well)/(avg change in control wells)] ⁇ 100). Typically, measurements were taken at 24 hour intervals and the period from 24-48 hours was used for %Inhibition measurements.
- the starting material for the isolation of the Mi antimicrobial protein was the basic fraction extracted from the mature seeds as described above in Example 1. This protein was further purified by cation exchange chromatography as shown in FIG. 1.
- FIGS. 1 a and 1 b Results of bioassays are included in FIGS. 1 a and 1 b where the elution gradient is shown as a solid line and the shaded bars represent %Inhibition.
- the FIG. 1 a assays were conducted without added Ca 2+ while 1 mM Ca 2+ was included in the FIG. 1 b assays.
- Fractionation yielded a number of unresolved peaks eluting between 0.05 and 2 M NaCl. A peak eluting at about 16 hours into the separation (fractions 90-92) showed significant antimicrobial activity.
- FIG. 2 shows the HPLC profile of purified fraction 92 from the cation-exchange separation shown in FIGS. 1 and 2. Protein elution was monitored at 214 nm. The acetonitrile gradient is shown by the straight line. Individual peaks were bioassayed for antimicrobial activity: the bars in FIG. 3 show the inhibition corresponding to 15 ⁇ g/ml of material from each of the fractions. The active protein elutes at approximately 27 min ( ⁇ 30% MeCN/0.1%TFA) and is called MiAMP2c.
- MiAMP2c was submitted for mass spectroscopic analysis. Approximately 1 ⁇ g of protein in solution was used for testing. Analysis showed the protein to have a molecular weight of 6216.8 Da ⁇ 2 Da. Additionally, the protein was subjected to reduction of disulfide bonds with dithiothreitol and alkylation with 4-vinylpyridine. The product of this reductionlalkylation was then submitted for mass spectroscopic analysis and was shown to have gained 427 mass units (i.e. molecular weight was increased by approximately 4 ⁇ 10 6 Da). The gain in mass indicated that four 4-vinylpyridine groups had reacted with the reduced protein, demonstrating that the protein contains a total of 4 cysteine residues. The cysteine content has also been subsequently confirmed through amino acid sequencing.
- the full amino acid sequence is RQRDP QQQYE QCQER CQRHE TEPRH MQTCQ QRCER RYEKE KRKQ KR and represents amino acids 118 to 164 of clone 3 from Example 9 (see FIG. 6 and SEQUENCE ID NO: 5).
- cysteine residues are in bold type and underlined to facilitate recognition of the spacing patterns.
- the protein mass will range from 6215.6 to 6219.6 Da. This is in close agreement with the mass of 6216.8 i 2 Da obtained by mass spectrometric analysis (Example 5).
- the measured mass closely approximates the predicted mass of MiAMP2c in a two-disulfide form as is expected to be the case.
- FIG. 5 shows a synthetic DNA sequence suitable for use in plant expression experiments.
- the arrow shows where translation is initiated and the triangular symbol indicates the point of cleavage of the signal peptide.
- FIG. 7 shows the predicted locations of ⁇ -helices, ⁇ -sheets and turns. The following symbols have been used in FIG. 7: C, coil (unstructured); H, alpha helix; E, ⁇ -sheet; and S, turn. Underlined residues are those which were predicted to exhibit an ⁇ -helical structure by at least 2 separate structure prediction methods; these are represented as helices in FIG. 8.
- the positively charged residues are the dark side chains outlined in black. Other dark side chains represent acidic residues.
- a proline residue (grey colour marked with a ‘P’) is located at the extreme left end of the molecule in the turn region. Solid black lines show where disulfide bonds connect the two helices. The dotted line shows where the two aromatic hydrophobic residues interact to add stability to the helix-turn-helix structure.
- This helix-turn-helix structure will be adopted by all MiAMP2 homologues containing the same cysteine spacing and residues with helix and turn-forming propensities.
- Other MiAMP2 fragment sequences can be superimposed onto the global structure shown in FIG. 8. The overall structure will remain essentially the same but the charge distribution will vary according to the sequences involved. In the case of MiAMP2b, the dotted line would represent an added disulfide bridge instead of a hydrophobic interaction.
- primer JPM17 sequence was 5′ CAG CAG CAG TAT GAG CAG TG 3′ and primer JPM20 degenerate sequence was 5′ TTT TTC GTA (T/T)C(T/G) (G/T)C(T/G) TTC GCA 3′ (SEQ ID NOS: 12 and 13).
- Primers JPM17 and JPM20 were used in PCR amplifications carried out for 30 cycles with 30 sec at 95° C., 1 min at 50° C., and 1 min at 72° C.
- PCR products with sizes close to those which were expected were directly sequenced (ABI PRISM Dye Terminator Cycle Sequencing Ready Reaction Kit from Perkin Elmer Corporation) after excising DNA bands from agarose gels and purifying them using a Qiagen DNA clean-up kit. Using this approach, it was possible to amplify a fragment of DNA of approximately 100 bp. Direct sequencing of this nucleotide fragment yielded the nucleotide sequence corresponding to a portion of the amino acid sequence of the antimicrobial protein MiAMP2c (amino acids 7-39 of FIG. 4).
- the partial nucleotide sequence obtained from the above-mentioned fragment excluding the primer sequences was 5′ TCA GAA GCG CTG CCA ACG GCG CGA GAC AGA GCC ACG ACA CAT GCA AAT TTG TCA ACA ACG C 3′ (corresponding to base pairs 264 to 324 in SEQ ID NO: 6).
- This sequence can be used for a variety of purposes including screening of cDNA and genomic libraries for clones of MiAMP2 homologues or design of specific primers for PCR amplification reactions.
- RNA from ground material was then purified using a Guanidine thiocyanate/Cesium chloride technique ( Current Protocols in Molecular Biology, supra). Using this method approximately 5 mg of total RNA was isolated. Messenger RNA was then purified from total RNA using a spun column mRNA purification kit (Pharmacia).
- a cDNA library was constructed in a lambda ZAP vector using a library kit from Stratagene. A total of 6 reactions were performed using 25 micrograms of messenger RNA. First and second strand cDNA synthesis was performed using MMLV Reverse transcriptase and DNA Polymerase I, respectively. After blunting the cDNA with Pfu DNA Polymerase, Eco RI linker adapters were ligated to the DNA. DNA was then kinased using T4 polynucleotide kinase and the DNA subsequently digested with Xho I restriction endonuclease. At this point cDNA material was fractionated according to size using a sephacryl-S500 column supplied with the kit. DNA was then ligated into the lambda ZAP vector. The vector containing ligated insert was then packaged into lambda phage (Gigapack III packaging extract from Stratagene).
- the library constructed above was then plated and screened in XL1-blue E. coli bacterial lawns growing in top agarose. Plaques containing individual clones were isolated by lifting onto Hybond N+ membranes (Amersham LIFE SCIENCE), hybridizing to a radiolabeled version of the genomic DNA fragment amplified above, imaging of the blot, and picking of possitive clones for the next round of screening. After secondary and tertiary screening, plaques were sufficiently isolated to allow picking of single clones. Several clones were obtained, and subsequently the pBK-CMV vector portion from the larger lambda vector was excised.
- clones 1 and 2 contained sequences differing from MiAMP2c by 2 residues and 3 residues, respectively, out of 47 amino acids total in the MiAMP2c sequence.
- the translation products of the full-length clones consist of a short signal peptide from residues 1 to 28, a hydrophilic region from residues 29 to ⁇ 246, and then two segments stretching from residues ⁇ 246 to 666 with a stretch of acidic residues separating them at positions 542-546.
- the hydrophilic region containing the sequence for MiAMP2c also contains 3 additional segments which are very similar to MiAMP2 (termed MiAMP2a, b and d). These 4 segments (found between residues 28 and ⁇ 246) are separated by stretches in which approximately four out of five residues are acidic (usually glutamic acid). These acidic stretches occur at positions 64-68, 111-115, 171-174, and 241-246 and appear to delineate processing sites for cleavage of the 666-residue preproprotein into smaller functional fragments (acidic stretches delineating cleavage sites are shown as bold characters in FIG. 6).
- MiAMP2-like segments of the protein contain 2 doublets of cysteine residues separated by 10-12 residues to give the following pattern C-X-X-X-C-(10-12X)-C-X-X-X-C where X is any amino acid, and C is cysteine. All four segments are expected to form helix-turn-helix motifs as decribed in Example 8 above. It is clear that the cysteines in these locations will form disulfide bridges that stabilize the structure of the proteins by holding the two helical portions together.
- the predicted helix-turn-helix motifs can be further stabilized in several ways.
- the first method of stabilization is exemplified in segments 1 and 3 (i.e., residues 29-63 and 118-170, respectively, of the 666-residue Macadamia vicilin-like protein). These segments is the are stabilized by a hydrophobic ring-stacking interaction between two aromatic residues (one on each ⁇ -helical segment); this is normally accomplished with tyrosine residues but phenylalanine is also used. As with the cysteine residues, the location of these aromatic residues in the predicted ⁇ -helical segments is critical if they are to offer stabilization to the helix-turn-helix structure.
- the aromatic residues are 2 and 3 residues removed from the cysteine doublets as shown here: Z-X-X-C-X-X-X-C-(10-12X)-C-X-X-X-C-X-X-Z where C is cysteine and Z is usually tyrosine but can be substituted with phenylalanine as is done in segment 1.
- the second way to stabilize the helix-turn-helix fragment is by using an added disulfide bridge as seen in fragment 2 (residues 71-110). This is accomplished by placing additional cysteine residues 2 and 3 residues removed from the cysteine doublets as shown here: nX-C-X-X-C-X-X-X-C-(10-12X)-C-X-X-X-C-X-X-C-nX.
- segment 4 does not contain the extra disulfide bridge or the hydrophobic ring-stacking stabilization, it is probably stabilized by means of weaker ionic and or hydrogen bonding interactions.
- PCR primers flanking the nucleotide region coding for MiAMP2c were engineered to contain restriction sites for Nde I and Bam HI (corresponding to the 5′ and 3′ ends of the coding region, respectively; Primer JPM31 sequence: 5′ A CAC CAT ATG CGA CAA CGT GAT CC 3′; Primer JPM32 sequence: 3′ C GTT GTT TTC TCT ATT CCT AGG GTT G 5′, SEQ ID NOS: 14 and 15). These primers were then used to amplify the coding region of MiAMP2c DNA.
- PCR product from this amplification was then digested with Nde I and Bam HI and ligated into a pET17b vector (Novagen/Studier, F. W. et al. [1986] J Mol. Biol. 189:113) with the coding region in-frame to produce the vector pET 17-MiAMP2c.
- the products were then digested with the appropriate restriction enzymes and ligated into the Nde I/Bam HI sites of a pET16b vector [Novagen] containing a His tag and a Factor Xa cleavage site (amino acid sequence MGHHH HHHHH HHSSG HIEGR HM, SEQ ID NO: 16).
- the protein products expressed from the pET16b vector is a fusion to the antimicrobial protein.
- the coding sequences for MiAMP2-like subunits from cocoa (FIG. 4, TcAMP1 and TcAMP2) were obtained from the published DNA sequence of the cocoa vicilin gene (Spencer, M. E. and Hodge R. [1992] Planta 186:567-576).
- Two MiAMP2-like fragments within the cocoa vicilin gene were located at the 5′ end (corresponding to the residues shown in FIG. 4), and two sets of complimentary oligonucleotides corresponding to the desired coding sequences were designed.
- the complimentary oligonucleotides (90 to ⁇ 100 bases) corresponding to each cocoa subunit contained a 20 bp overlap and also contained the Nde I and Bam HI restriction endonuclease cut sites.
- TcAMP2 forward oligo 5′ GGGAATTCCA TATGCTTCAA AGGCAATACC AGCAATGTCA AGGGCGTTGT CAAGAGCAAC AACAGGGGCA GAGAGAGCAG CAGCAGTGCC AGAGAAAATG C 3′; TcAMP2 reverse oligo 5′ GTGTGGATCC CTAGCTCCTA TTTTTTTTGT GATTATGGTA ATTCTCGTGC TCGCCTCTCT CTTGTTCCTT ATATTGCTCC CAGCATTTTC TCTGGCACTG CT 3′.
- oligonucleotide sets were added to individual PCR amplification reactions in order make individual PCR fragments containing the desired coding region. Since initial PCR amplifications gave fuzzy bands, reamplification of the original products was carried out using new 20mer primers (complimentary to the 5′ends of the forward and reverse oligonucleotides shown above) designed to amplify the entire coding region of the cocoa subunits. Once amplified, the PCR products were restriction digested with the appropriate enzymes and ligated into the vector pET16b as above. This procedure was carried out for both cocoa fragments with similarities to MiAMP2c (shown in FIG. 4).
- MiAMP2c homologues except MiAMP2c which was expressed in pET17b
- pET16b vector containing the Histidine tag While induction of the MiAMP2c culture proceeded as above, the rest of the purification was somewhat different. In this case, MiAMP2c-expressing cells were harvested by centrifugation but were then resuspended in phosphate buffer (100 mM, pH 7.0 containing 10 mM EDTA and 1 mM PMSF) and broken open using a French press instrument. Cellular debris containing MiAMP2c inclusion bodies was solubilized using a 6 M Guanidine-HCl, 10 mM MES pH 6.0 buffer.
- FIG. 9 shows the SDS-PAGE gel analysis of the various purification stages obtained following induction with IPTG and subsequent purification of expressed proteins.
- Samples analysed during the TcAMP1 purification were are as follows: lane 1, molecular weight markers; lane 2, Ni-NTA non-binding fraction; lane 3, rinse of Ni-NTA resin with pH 8 urea; lane 4, rinse of Ni-NTA resin with pH 6.3 urea; lane 5, elution of TcAMP1 with pH 4.5 urea; and lane 6, second elution of TcAMP1 with pH 4.5 urea.
- TcAMP2 was purified in a similar manner and was also subjected to reverse-phase HPLC to further purify the fraction eluting from the Ni-NTA resin.
- FIG. 10 shows the reverse phase purification of cocoa subunit number 2 (TcAMP2).
- SDS-PAGE gel analysis of the MiAMP2a, b, and d fragment purification is shown in the second panel of FIG. 9.
- Lane contents are as follows: lane 1, molecular weight markers; lane 2, MiAMP2a pre-induced cellular extractp; lane 3, MiAMP2a IPTG induced cellular extract; lane 4, MiAMP2a Ni-NTA non-binding fraction; lane 5, MiAMP2a elution from Ni-NTA; lane 6, MiAMP2b pre-induced cellular extract; lane 7, MiAMP2b IPTG induced cellular extract; lane 8, MiAMP2b Ni-NTA non-binding fraction; lane 9, MiAMP2b elution from Ni-NTA; lane 10, MiAMP2d pre-induced cellular extract; lane 11, MiAMP2d IPTG induced cellular extract; lane 12, MiAMP2d Ni-NTA non-binding fraction; and lane 13, MiAMP2d elution from Ni-NTA.
- MiAMP2c, and 5 homologues i.e., MiAMP2a, MiAMP2b, MiAMP2d, TcAMP1 and TcAMP2
- MiAMP2a, MiAMP2b, MiAMP2d, TcAMP1 and TcAMP2 were all expressed, purified and tested for antimicrobial activity.
- the approach taken above can be applied to all of the antimicrobial fragments described in FIG. 4. Purified fragments can then be tested for specific inhibition agains microbial pathogens of interest.
- FIG. 11 shows that various other species contain immunologically-related proteins of similar size to MiAMP2c.
- Lanes 1-15 contain the extracts from the following species: 1) Stenocarpus sinuatus, 2) Stenocarpus sinuatus ( ⁇ fraction (1/10) ⁇ loading), 3) Restio tremulus, 4) Mesomalaena tetragona, 5) Nitraria billardieri, 6) Petrophile canescens, 7) Synaphae acutiloba, 8) Dryandra formosa, 9) Lambertia inermis, 10) Stirlingia latifolia, 11) Xylomelum angustifolium, 12) Conospermum bracteosum, 13) Conospermum triplinernium, 14) Molecular weight marker, 15) Macacamia integrifolia pure MiAMP2c.
- Lanes 1-13 contain a variety of species, some of which show the presence of antigenically related proteins of a similar size to MiAMP2c. Other bands exhibiting higher molecular weights probably represent the larger precursor seed storage proteins from which the antimicrobial proteins are derived. Antigenically-related proteins can be seen in lanes 1, 2, 4, 6, 7, 8, 9, and 11-13.
- Bioassays were also performed using crude extracts from various Proteaceae species. Specifically, extracts from Banksia robur, Banksia canei, Hakea gibbosa, Stenocarpus sinuatus, and Stirlingia latifolia have all been shown to exhibit antimicrobial activity. This activity may derive from MiAMP2 homologues since these species are related to Macadamia.
- Stenocarpus sinuatis was chosen for a large scale fractionation experiment in an attempt to isolate MiAMP2c homologues.
- Five kg of S. sinuatus seed was frozen in liquid nitrogen and ground in a food processor (Big Oscaar Sunbeam).
- the ground seed was immediately placed into 12 L of 50 mM H 2 SO 4 extraction buffer and extracted at 4° C. for 1 hour with stirring.
- the slurry was then centrifuged for 20 min at 10,000 g to remove particulate matter.
- the supernatant was then adjusted to pH 9 using a 50 mM ammonia solution.
- PMSF and EDTA were added to final concentrations of 1 and 10 mM respectively.
- the crude protein extract was applied to an anion exchange column (Amberlite IRA-938, Rohm and Haas) (3 cm ⁇ 90 cm) equilibrated with 50 mM NH 4 Ac pH 9.0 at a flow rate of 40 ml/min.
- the unbound protein comprising the basic protein fraction was collected and used in the subsequent purification steps.
- the basic protein fraction was adjusted to pH 5.5 with acetic acid and then applied at 10 ml/minute over 12 h to a SP-Sepharose Fast Flow (Pharmacia) Column (5 cm ⁇ 60 cm) pre-equilibrated with 25 mM ammonium acetate. The column was then washed for 3.5 h with 25 mM Acetate pH 5.5. Elution of bound proteins was achieved by applying a linear gradient of NH 4 Ac from 25 mM to 2.0 M (pH 5.5) at 10 ml/min over 10 h. Absorbance of the eluate was observed at 280 nm and 100 ml fractions collected (see FIG. 12).
- Cation-exchange fractions that cross-reacted with the antiserum were then further purified by reverse phase chromatography.
- Bound proteins were eluted with a linear gradient from 100%A to 100%B (5% H 2 O, 95% acetonitrile, 0.08% TFA). The absorbance of the eluted proteins was monitored at 214 nm and 280 nm.
- Peptide one is comprised of 22 amino acids from 118 to 139 in the amino acid sequence of clone 3 (sequence: RQRDP QQQAE QAQKR AQRRE TE, SEQUENCE ID NO: 9).
- Peptide 2 is 25 amino acids in length and runs from 140 to 164 in clone 3 (sequence: PRHMQ IAQQR AERRA EKEKR KQQKR, SEQ ID NO: 10).
- Peptides 1 and 2 are labeled MiAMP2c pep1 and MiAMP2c pep2 respectively. These peptides were resuspended in Milli-Q water and bioassayed against a number of fungi.
- peptide 2 has inhibitory activity against a variety of fungi whereas peptide 1 exhibited little or no activity. Mixtures of peptide 1 and peptide 2 exhibit similar levels of activity as seen with peptide 2 alone indicating that only peptide 2 is exhibiting activity. The fact that peptide 2 exhibits antimicrobial activity in the absence of the helix-turn-helix structure exhibited by MiAMP2c reveals that the helix-turn-helix structure is not absolutely necessary for the peptides to retain activity.
- peptide 2 did not exhibit the same degree of activity on a molar basis as MiAMP2c (whole fragment) indicating that the helix-turn-helix structure is important for maximal expression of antimicrobial activity by the fragments involved. It is also expected that the helix-turn-helix structure will confer greater stability to the MiAMP2 homologues, thus rendering these proteins less susceptible to proteolytic cleavage and other forms of degradation. Greater stability would lead to maintaining antimicrobial activity over a longer period of time.
- MiAMP2c and each of the various MiAMP2 homologues were tested against a variety of fungi as concentrations ranging from 2 to 50 ⁇ g/ml.
- Table 1 shows the IC 50 value of pure MiAMP2c against various fungi and bacteria.
- the “>50” indicates that 50% inhibition of the fungus was not achieved at 50 ⁇ g/ml which was the highest concentration tested.
- the abbreviation “ND” indicates that the test was not performed or that results could not be interpreted.
- the antimicrobial activity of MiAMP2c was also tested in the presence bf 1 mM Ca 2+ in the test medium and the IC 50 values for these tests are given in the right-hand column.
- Table 2 shows the antimicrobial activity of various homologues and fragments of MiAMP2c.
- Ab Alternaria brassicola
- Cp Ceratocystis paradoxa
- Foc Fusarium oxysporum
- Lm Leptosphaeria maculans
- Ss Sclerotinia sclerotiorum
- Vd Verticillium dahlias.
- the “>50” indicates that concentrations higher than 50 ⁇ g/ml were not tested so that an IC 50 value could not be established.
- a blank space indicates that the test was not performed or that results could not be interpreted.
- TcAMP1 and 2 used for the results presented in Table 2 were derived from cocoa vicilin (Examples 10 and 11). SsAMP1 and 2 show reactivity with MiAMP2c antibodies and also exhibit antimicrobial activity as seen in the table below.
- the versions of MiAMP2a, b and d as well as TcAMP1 and TcAMP2 tested in the bioassays all contain a His tag fusion resulting from expression in the vector pET16b.
- MiAMP2c pep1 and 2 are the N and C terminal regions, respectively, of MiAMP2c antimicrobial peptide as specified in Example 14 above.
- concentration value listed for ‘MiAMP2c pep1+2’ is the concentration of each individual peptide in the mixture.
- MiAMP2c pep1 and pep2 are both about 1 ⁇ 2 the size of MiAMP2c; comparisons of the activity of these peptides with the MiAMP2c protein should, therefore, be made on a molar basis rather than on a strict ⁇ g/ml concentration basis. Peptides were only tested in media A which did not contain added Ca 2+ .
- TcAMP1 and 2 sequences are readily available in the public data bases, no antimicrobial activity had ever been assigned to them. These sequences were derived from much larger proteins involved in seed storage functions. The inventors have thus described a completely new activity for a small portion of the overall cocoa vicilin molecules. The activity of cotton fragments 1, 2, and 3 has been exemplified by other authors (Chung, R. P. T. et al. [1997 ] Plant Science 127:1-16).
- the expression vector pPCV91-MiAMP2c contains the full coding region of the MiAMP2c (Example 7) DNA flanked at it 5′ end by the strong constitutive promoter of 35S RNA from the cauliflower mosaic virus (pCaMV35S) (Odel et al., [1985] Nature 313: 810-812) with a quadruple-repeat enhancer element (e-35S) to allow for high transcriptional activity (Kay et al. [1987] Science 236:1299-1302).
- the coding region of MiAMP2c DNA is flanked at its 3′ end by the polyadenylation sequence of 35S RNA of the cauliflower mosaic virus (pA35S).
- the plasmid backbone of this vector is the plasmid pPCV91 (Walden, R. et al. [1990] Methods Mol. Cell. Biol. 1:175-194).
- the plasmid also contains other elements useful for plant transformation such as an ampicillin resistance gene (bla) and a hygromycin resistance gene (hph) driven by the nos promoter (pnos). These and other features allow for selection in various cloning and transformation procedures.
- the plasmid pPCV91-MiAMP2c was constructed as follows: A cloned fragment encoding MiAMP2c (Example 7) was digested using restriction enzymes to release the MiAMP2c gene fragment containing a synthetic leader sequence.
- the binary vector pPCV91 was digested with the restriction enzyme Bam HI. Both the MiAMP2c DNA fragment containing and the binary vector were ligated using T4 DNA ligase to produce pPCV9 1-MiAMP2c binary vector for plant transformation (FIG. 12).
- MiAMP2c can be expressed in plants. Not only can individual homologues be expressed, but they may be expressed in combination with other proteins as fusion proteins or as portions of larger precursor proteins. For example, it is possible to express the N-terminal region of MiAMP2 clone 1 (amino acids 1 to ⁇ 246) which contains a signal peptide and the hydrophilic region containing four antimicrobial segments. Transgenic plants can then be assessed to examine whether the individual fragments are being processed into the expected fragments by the processing machinery already present in the plant cells. It is also possible to express the entire MiAMP2 clone 1 (amino acids 1 to 666) and to examine the processing of the entire protein when expressed in transgenic plants.
- amino acids 1 to ⁇ 246 amino acids 1 to ⁇ 246
- Homologous regions from other sequences can also be used in multiple combinations with, for example, ten (10) or more MiAMP2-like fragments expressed as one large fusion protein with acidic cleavage sites located as proper locations between each of the fragments. As well as linking MiAMP2 fragments together, it would also be possible to link MiAMP2 fragments to other useful proteins for expression in plants.
- Tobacco transformation was carried out using leaf discs of Nicotiana tabacum based on the method of Horsch et al. ( Science 227:1229-1231 [1985]) and co-culturing strains containing pPCV91-MiAMP2c. After co-cultivation of Agrobacterium and tobacco leaf disks, transgenic plants (transformed with pPCV91-MiAMP2c) were regenerated on media containing 50 ⁇ g/ml hygromycin and 500 ⁇ g/ml Cefotaxime. These transgenic plants were analysed for expression of the newly-introduced genes using standard western blotting techniques (FIG. 15).
- FIG. 15 standard western blotting techniques
- Lane 15 shows a western blot of extracts from trangenic tobacco carrying the construct for MiAMP2c from example 16.
- Lane 1 contains pure MiAMP2c as a standard
- lanes 2 and 3 contain extracts from transgenic plants canying the pPCV9 1-MiAMP2c construct. As can be see in the figure, faint bands are present at approximately the correct molecular weight, indicating that the transgenic plants appear to be expressing the MiAMP2c protein. Plants capable of constitutive expression of the introduced genes may be selected and self-pollinated to give seed. Fl seedlings of the transgenic plants may be further analysed.
Landscapes
- Health & Medical Sciences (AREA)
- Life Sciences & Earth Sciences (AREA)
- Engineering & Computer Science (AREA)
- General Health & Medical Sciences (AREA)
- Genetics & Genomics (AREA)
- Chemical & Material Sciences (AREA)
- Biotechnology (AREA)
- Wood Science & Technology (AREA)
- Organic Chemistry (AREA)
- Zoology (AREA)
- Natural Medicines & Medicinal Plants (AREA)
- Plant Pathology (AREA)
- Microbiology (AREA)
- Molecular Biology (AREA)
- Biomedical Technology (AREA)
- Environmental Sciences (AREA)
- Agronomy & Crop Science (AREA)
- Dentistry (AREA)
- Mycology (AREA)
- General Engineering & Computer Science (AREA)
- Bioinformatics & Cheminformatics (AREA)
- Biophysics (AREA)
- Biochemistry (AREA)
- Medicinal Chemistry (AREA)
- Physics & Mathematics (AREA)
- Animal Behavior & Ethology (AREA)
- General Chemical & Material Sciences (AREA)
- Gastroenterology & Hepatology (AREA)
- Botany (AREA)
- Veterinary Medicine (AREA)
- Public Health (AREA)
- Pharmacology & Pharmacy (AREA)
- Nuclear Medicine, Radiotherapy & Molecular Imaging (AREA)
- Cell Biology (AREA)
- Proteomics, Peptides & Aminoacids (AREA)
- Chemical Kinetics & Catalysis (AREA)
- Oncology (AREA)
- Communicable Diseases (AREA)
- Peptides Or Proteins (AREA)
- Medicines That Contain Protein Lipid Enzymes And Other Medicines (AREA)
Abstract
A new family of antimicrobial proteins is described. Prototype proteins can be isolated from Macadamia integrifolia as well as other plant species. DNA encoding the protein is also described as well as DNA constructs which can be used to express the antimicrobial protein or to introduce the antimicrobial protein into a plant. Compositions comprising the antimicrobial protein or the antimicrobial protein per se can be administered to plants or mammalian animals to combat microbial infestation.
Description
- This invention relates to isolated proteins which exert inhibitory activity on the growth of fungi and bacteria, which fungi and bacteria include some microbial pathogens of plants and animals. The invention also relates to recombinant genes which include sequences encoding the proteins, the expression products of which recombinant genes can contribute to plant cells or cells of other organism's defence against invasion by microbial pathogens. The invention further relates to the use of the proteins and/or genes encoding the proteins for the control of microbes in human and veterinary clinical conditions.
- Microbial diseases of plants are a significant problem to the agricultural and horticultural industries. Plant diseases in general cause millions of tonnes of crop losses annually with fungal and bacterial diseases responsible for significant portions of these losses. One possible way of combating fungal and bacterial diseases is to provide transgenic plants capable of expressing a protein or proteins which in some way increase the resistance of the plant to pathogen attack. A simple strategy is to first identify a protein with antimicrobial activity in vitro, to clone or synthesise the DNA sequence encoding the protein, to make a chimaeric gene construct for efficient expression of the protein in plants, to transfer this gene to transgenic plants and to assess the effect of the introduced gene on resistance to microbial pathogens by comparison with control plants.
- The first and most important step in the strategy for disease control described above is to identify, characterise and describe a protein with strong antimicrobial activity. In recent years, many different plant proteins with antimicrobial and/or antifungal activity have been identified and described. These proteins have been categorised into several classes according to either their presumed mode of action and/or their amino acid sequence homologies. These classes include the following: chitinases (Roberts, W. K. et al. [1986 ] Biochim. Biophys. Acta 880:161-170); β-1,3-glucanases (Manners, J. D. et al. [1973] Phytochemistry 12:547-553); thionins (Bolmann, H. et al. [1988] EMBO J. 7:1559-1565 and Fernadez de Caleya, R. et al. [1972] Appl. Microbiol. 23:998-1000); permatins (Roberts, W. K. et al. [1990] J. Gen. Microbiol. 136:1771-1778 and Vigers, A. J. et al [1991] Mol. Plant-Microbe Interact. 4:315-323); ribosome-inactivating proteins (Roberts, W. K. et al. [1986] Biochini. Biophys. Adcta 880:161-170 and Leah, R. et al. [1991] J. Biol. Chem. 266:1564-1573); plant defensins (Terras, F. R. G. et al. [1995] The Plant Cell 7:573-588); chitin binding proteins (De Bolle, M. F. C. et al. [1992] Plant Mol. Biol. 22:1187-1190 and Van Parijs, J. et al. [1991] Planta 183:258-264); thaumatin-like, or osmotin-like proteins (Woloshuk, C. P. et al. [1991] The Plant Cell 3:619-628 and Hejgaard, J. [1991] FEBS Letts. 291:127-131); PR1-type proteins (Niderman, T. et al. [1995] Plant Physiol. 108:17-27.) and the non-specific lipid transfer proteins (Terras, F. R. G. et al. [1992] Plant Physiol. 100:1055-1058 and Molina, A. et al. [1993] FEBS Letts. 3166:119-122). Another class of antimicrobial proteins from plants is the knottin or knottin-like antimicrobial proteins (Cammue, B.P. A. et al. [1992] J. Biol. Chem. 67:2228-2233; Broekaert W. F. et al. (1997) Crit. Rev. in Plant Sci. 16(3):297-323). A class of antimicrobial proteins termed 4-cysteine proteins has also been reported in the literature which class includes Maize Basic Protein (MBP-1) (Duvick, J. P. et al. [1992] J. Biol. Chem. 267:18114-18120). A novel antimicrobial protein which does not fit into any previously described class of antimicrobial proteins has also been isolated from the seeds of Macadamia integrifolia termed MiAMP1 (Marcus, J. P. et al. [1997] Eur. J. Biochem. 244:743-749). In addition, plants are not the sole source of antimicrobial proteins and there are many reports of the isolation of antimicrobial proteins from animal and microbial cells (reviewed in Gabay, J. E. [1994] Science 264:373-374 and in “Antimicrobial peptides” [1994] CIBA Foundation Symposium 186, John Wiley and Sons Publ., Chichester, UK).
- There is evidence that the ectopic expression of genes encoding proteins that have in vitro antimicrobial activity in transgenic plants can result in increased resistance to microbial pathogens. Examples of this engineered resistance include transgenic plants expressing genes encoding: a plant chitinase, either alone (Broglie, K. et al. [1991] Science 254:1194-1197) or in combination with a 1,3-glucanase (Van den Elzen, P. J. M. et al. [1993] Phil. Trans. Roy. Soc. 342:271-278); a plant defensin (Terras, F. R. G. et al. [1995] The Plant Cell 7:573-588); an osmotin-like protein (Liu, D. et al. [1994] Proc. Natl. Acad. Sci. USA 91:1888-1892); a PR1-class protein (Alexander, D. et al. [1993] Proc. Natl. Acad. Sci. USA 90:7327-7331) and a ribosome-inactivating protein (Logemann, J. et al. [1992] Bio/Technology 10:305-308).
- Although the potential use of antimicrobial proteins for engineering disease resistance in transgenic plants has been described extensively, there are other applications which are worthy of mention. Firstly, highly potent antimicrobial proteins can be used for the control of plant disease by direct application (De Bolle, M. F. C. et al. [1993] in Mechanisms of Plant Defence Responses, B. Fritig and M. Legrand eds., Kluwer Acad. Publ., Dordrecht, NL, pp. 433-436). In addition, antimicrobial peptides have potential therapeutic applications in human and veterinary medicine. Although this has not been described for peptides of plant origin it is being actively explored with peptides from animals and has reached clinical trials (Jacob, L. and Zasloff, M. [1994] in “Antimicrobial Peptides”, CIBA Foundation Symposium 186, John Wiley and Sons Publ., Chichester, UK, pp. 197-223).
- Antimicrobial proteins exhibit a variety of three-dimensional structures which will determine in large part the activity which they manifest. Many of the global structures exhibited by these proteins have been determined (Broekaert W.F. et a. (1997) Crit. Rev. in Plant Sci. 16(3):297-323). A large factor in determining the stability of these proteins is the presence of disulfide bridges between various cysteines located in a helical and β-sheet regions. Many peptides with toxic activity such as conotoxin are well known to be stabilized by disulfide bridges (see for example Hill, J. M. et al. (1996) Biochemistry 35(27): 8824-8835). In the case of the conotoxin referenced above, a compact structure is formed consisting of a helix, a small -hairpin, a cis-hydroxyproline, and several turns. The molecule is stabilized by three disulfide bonds, two of which connect the α-helix and the β-sheet, forming a solid structural core. Interestingly, eight arginine and lysine side chains in this molecule project into the solvent in a radial orientation relative to the core of the molecule. These cationic side chains form potential sites of interaction with anionic sites on pathogen membranes (Hill, J. M. et al. supra).
- The invention described herein constitutes previously undiscovered and thus novel proteins with antimicrobial activity. These proteins can be isolated from Macadamia integrifolia (Mi) seeds or from cotton or cocoa seeds. In addition, protein fragments which are antifungal can be derived from larger seed storage proteins containing regions of substantial similarity to the antimicrobial proteins from macadamia described here. Examples of seed storage proteins which contain regions similar to the proteins which have been purified can be seen in FIG. 4. Macadamia integrifolia belongs to the family Proteaceae. M. integrifolia, also known as Bauple Nut or Queensland Nut, is considered by some to be the world's best edible nut. Cotton (Gossypium hirsutum) belongs to the family Malvaceae and is cultivated extensively for its fiber. Cocoa (Threobroma cacao) belongs to the family Sterculiaceae and is used around the world for a wide variety of cocoa products.
- The fact that both the macadamia and cocoa antimicrobial proteins are found in edible portions of these plants makes these peptides attractive for use in genetic engineering for disease resistance since trangenic plants expressing these proteins are unlikely to show added toxicity. Proteins may also be safe for human and veterinary use.
- According to a first embodiment of the invention, there is provided a protein fragment having antimicrobial activity, wherein said protein fragment is selected from:
- (i) a polypeptide having an amino acid sequence selected from:
-
residues 29 to 73 of SEQ ID NO: 1 - residues 74 to 116 of SEQ ID NO: 1
-
residues 117 to 185 of SEQ ID NO: 1 - residues 186 to 248 of SEQ ID NO: 1
-
residues 29 to 73 of SEQ ID NO: 3 - residues 74 to 186 of SEQ ID NO: 3
- residues 167 to 185 of SEQ ID NO: 3
- residues 186 to 248 of SEQ ID NO: 3
-
residues 1 to 32 of SEQ ID NO: 5 -
residues 33 to 75 of SEQ ID NO: 5 - residues 76 to 144 of SEQ ID NO: 5
- residues 145 to 210 of SEQ ID NO: 5
-
residues 34 to 80 of SEQ ID NO: 7 - residues 81 to 140 of SEQ ID NO: 7
-
residues 33 to 79 of SEQ ID NO: 8 -
residues 80 to 11 9 of SEQ ID NO: 8 -
residues 120 to 161 of SEQ ID NO: 8 -
residues 32 to 91 of SEQ ID NO: 21 - residues 25 to 84 of SEQ ID NO: 22
-
residues 29 to 94 of SEQ ID NO: 24 -
residues 31 to 85 of SEQ ID NO: 25 -
residues 1 to 23 of SEQ ID NO: 26 -
residues 1 to 17 of SEQ ID NO: 27 -
residues 1 to 28 of SEQ ID NO: 28; - (ii) a homologue of (i);
- (iii) a polypeptide containing a relative cysteine spacing of C-2X-C-3X-C-(10-12)X-C-3X-C-3X-C wherein X is any amino acid residue, and C is cysteine;
- (iv) a polypeptide containing a relative cysteine and tyrosine/phenylalanine spacing of Z-2X-C-3X-C-(10-12)X-C-3X-C-3X-Z wherein X is any amino acid residue, and C is cysteine, and Z is tyrosine or phenylalanine;
- (v) a polypeptide containing a relative cysteine spacing of C-3X-C-(10-12)X-C-3X-C wherein X is any amino acid residue, and C is cysteine;
- (vi) a polypeptide with substantially the same spacing of positively charged residues relative to the spacing of cysteine residues as (i); and
- (vii) a fragment of the polypeptide of any one of (i) to (vi) which has substantially the same antimicrobial activity as (i).
- According to a second embodiment of the invention, there is provided a protein containing at least one polypeptide fragment according to the first embodiment, wherein said polypeptide fragment has a sequence selected from within a sequence comprising SEQ ID NO: 1, SEQ ID NO: 3 or SEQ ID NO: 5.
- According to a third embodiment of the invention, there is provided a protein having a sequence selected from SEQ ID NO: 1, SEQ ID NO: 3 or SEQ ID NO: 5.
- According to a fourth embodiment of the invention, there is provided an isolated or synthetic DNA encoding a protein according to the first embodiment According to a fifth embodiment of the invention, there is provided a DNA construct which includes a DNA according to the fourth embodiment operatively linked to elements for the expression of said encoded protein.
- According to a sixth embodiment of the invention, there is provided a transgenic plant harbouring a DNA construct according to the fifth embodiment.
- According to a seventh embodiment of the invention, there is provided reproductive material of a transgenic plant according to the sixth embodiment.
- According to an eighth embodiment of the invention, there is provided a composition comprising an antimicrobial protein according to the first embodiment together with an agriculturally-acceptable carrier diluent or excipient.
- According to a ninth embodiment of the invention, there is provided a composition comprising an antimicrobial protein according to the first embodiment together with an pharmaceutically-acceptable carrier diluent or excipient.
- According to a tenth embodiment of the invention, there is provided a method of controlling microbial infestation of a plant, the method comprising:
- i) treating said plant with an antimicrobial protein according to the first embodiment or a composition according to the eighth embodiment; or
- ii) introducing a DNA construct according to the fifth embodiment into said plant.
- According to an eleventh embodiment of the invention, there is provided a method of controlling microbial infestation of a mammalian animal, the method comprising treating the animal with an antimicrobial protein according to the first embodiment or a composition according to the ninth embodiment.
- According to a twelfth embodiment of the invention, there is provided a method of preparing an antimicrobial protein, which method comprises the steps of:
- a) obtaining or designing an amino acid sequence which forms a helix-turn-helix structure;
- b) replacing individual residues to achieve substantially the same distribution of positively charged residues and cysteine residues as in one or more of the amino acid sequences shown in FIG. 4;
- c) synthesising a protein comprising said amino acid sequence chemically or by recombinant DNA techniques in liquid culture; and
- d) if necessary, forming disulphide linkages between said cysteine residues.
- Other embodiments of the invention include methods for producing antimicrobial protein.
- FIG. 1 shows the results of cation-exchange chromatography of the basic protein fraction of a Macadamia integrifolia extract with the results of a bioassay for antimicrobial activity shown for fractions in the region of MiAMP2c elution.
- FIG. 2 shows the results of including 1 mM Ca 2+ in a parallel bioassay of fractions from the cation-exchange separation.
- FIG. 3 shows a reverse-phase HPLC profile of highly inhibitory fractions containing MiAMP2c from the cation-exchange separation in FIGS. 1 and 2 together with % growth inhibition exhibited by the HPLC fractions.
- FIG. 4 shows the amino acid sequences of MiAMP2a, b, c and d and protein fragments derived from other seed storage proteins which contain regions of homology to the MiAMP2 series of antimicrobial proteins.
- FIG. 5 shows an example of a synthetic nucleotide sequence which can be used for the expression and secretion of MiAMP2c in transgenic plants.
- FIG. 6 shows the alignment of clones 1-3 from macadamia containing MiAMP2a, b, c and d subunits together with sequences from cocoa and cotton vicilin seed storage proteins which exhibit significant homology to the macadamia clones.
- FIG. 7 displays a series of secondary structure predictions for MiAMP2c.
- FIG. 8 shows a three-dimensional model of the MiAMP2c protein.
- FIG. 9 shows stained SDS-PAGE gels of protein fractions at various stages in the expression and purification of TcAMP1 ( Theobroma cacao subunit 1), MiAMP2a, MiAMP2b, MiAMP2c and MiAMP2d expressed in E. coli liquid culture.
- FIG. 10 shows the reverse-phase HPLC purification of cocoa subunit 2 (TcAMP2) after the initial purification step using Ni-NTA media.
- FIG. 11 shows a western blot of crude protein extracts from various plant species using rabbit antiserum raised to MiAMP2c.
- FIG. 12 shows a cation-exchange fractionation of the Stenocarpus sinuatus basic protein fraction along with the accompanying western blot which shows the presence of immunologically-related proteins in a range of fractions.
- FIG. 13 shows a reverse-phase HPLC separation of Stenocarpus sinuatus cation-exchange fractions which had previously reacted with MiAMP2c antibodies (see FIG. 14). A western blot is also presented which reveals the presence of putative MiAMP2c homologues in individual HPLC fractions.
- FIG. 14 is a map of the binary vector pPCV91-MiAMP2c as an example of a vector that can be used to express these antimicrobial proteins in transgenic plants.
- FIG. 15 shows a western blot to detect MiAMP2c expressed in transgenic tobacco plants.
- The following abbreviations are used hereafter:
EDTA ethylenediaminetetraacetic acid IPTG Isopropyl-β-D-thiogalactopyranoside MeCN methyl cyanide (acetonitrile) Mi Macadamia integrifolia MiAMP2 Macadamia integrifolia antimicrobial protein series number 2 Ni-NTA Nickel-nitrilotriacetic acid chromatography media ND not determined PCR polymerase chain reaction PMSF phenylmethylsulphonyl fluoride SDS-PAGE sodium-dodecylsulphate polyacrylamide gel electrophoresis TFA trifluoroacetate - The term homologue is used herein to denote any polypeptide having substantial similarity in composition and sequence to the polypeptide used as the reference. The homologue of a reference polypeptide will contain key elements such as cysteine or other residues spaced at identical intervals such that a substantially similar three-dimensional global structure is adopted by the homologue as compared to the reference. The homologue will also exhibit substantially the same antimicrobial activity as the reference protein.
- The present inventors have identified a new class of proteins with antimicrobial activity. Prototype proteins can be isolated from seeds of Macadamia integrifolia. The invention thus provides antimicrobial proteins per se and also DNA sequences encoding these antimicrobial proteins.
- The invention also provides amino acid sequences of proteins which are homologous to the prototype antimicrobial proteins from Macadamia integrifolia. Thus, in addition to the antimicrobial proteins from Macadamia, this invention also provides amino acid sequences of homologues from other species which have hitherto been unrecognized as having antimicrobial activity.
- While the first antimicrobial protein in the present series was isolated directly from Macadamia integrifolia, additional antimicrobial proteins were identified through cloning efforts, homology searches and subsequent antimicrobial testing of the encoded proteins after expression in and purification from liquid culture. After the first protein from this series was purified from macadamia and termed MiAMP2, clones were obtained which encoded a preproprotein containing MiAMP2. This large protein (666 amino acids), represented by several almost identical clones, contained four adjacent regions with significant similarity to the purified antimicrobial protein fragment (MiAMP2) which itself was found to lie within region three in the cloned nucleotide sequence; hence the purified antimicrobial protein is termed MiAMP2c. Other fragments contained in the 666-amino-acid clone are termed MiAMP2a, b and d as per their locations in the cloned nucleotide sequence. Several other sequences with significant homology to the MiAMP2a, b, c, and d protein fragments were then identifed in the Entrez data base. These homologous sequences were contained within larger seed storage proteins from cotton and cocoa which sequences had not been previously described as containing antimicrobial protein sequences or as exhibiting antimicrobial activity. Fragments of larger seed storage proteins containing sequences homologous to MiAMP2c were tested and are here demonstrated to exhibit antimicrobial activity. Thus, the inventors have established a process for obtaining antimicrobial protein fragments from larger seed storage proteins. In the light of these findings, it is evident that fragments of other seed storage proteins containing sequences similar to the proteins described will also exhibit antimicrobial activity.
- In particular, the 47-amino-acid TcAMP1 (for Theobroma cacao antimicrobial protein 1) and the 60-amino-acid TcAMP2 sequences were derived from a cocoa vicilin seed storage protein gene sequence (which contains 525 amino acids) (Spencer, M. E. and Hodge R. [1992] Planta 186:567-576). These derived fragments were then expressed in liquid culture. Cocoa vicilin fragments thus expressed and subsequently purified (Examples 10 and 11), were shown to be antimicrobial (Example 15). This is the first report that fragments of the cocoa vicilin protein possess antimicrobial activity. Pools of sequences containing fragments homologous to the MiAMP2c apparently released from cotton vicilin seed storage protein have been shown to possess antimicrobial activity (Chung, R. P. T. et al. [1997] Plant Science 127:1-16). This finding is clearly embodied in sequences disclosed in this application.
- In addition to showing that cocoa-vicilin-derived fragments exhibit antimicrobial activity, there is herein described additional proteins which exhibit antimicrobial activity. For example, there is described below proteins from Stenocarpus sinuatus which are of similar size to MiAMP2 subunits, react with MiAMP2c antiserum, and contain sequences homologous to MiAMP2 proteins (see FIG. 4). Based on the evidence provided herein, sequences homologous to the MiAMP2c subunit (i.e., MiAMP2a, b, d; TcAMP1; TcAMP2; and
1, 2 and 3—see FIG. 4) constitute proteins which contain the fragment with antimicrobial activity. The antimicrobial activity of MiAMP2 fragments from macadamia, and the TcAMP1 and 2 fragments from cocoa, is exemplified below. R. P. T. Chung et al. (Plant Science 127:1-16 [1997]) have demonstrated that the cotton fragments exhibit antimicrobial activity. Other antimicrobial proteins can also be derived from seed storage proteins such as peanut allergen Ara h (Burks, A. W. et al. [1995] J. Clin. Invest. 96 (4), 1715-1721), maize globulin (Belanger, F. C. and Kriz, A. L.[1991] Genetics 129 (3), 863-872), barley globulin (Heck, G. R. et al. [1993] Mol. Gen. Genet. 239 (1-2), 209-218), and soybean. conglycinin (Sebastiani, F. L. et al. [1990] Plant Mol. Biol. 15 (1), 197-201), all of which contain the same key elements which are present in the sequences which are here shown to exhibit antimicrobial activity.cotton fragments - The proteins which contain regions of sequence homologous to MiAMP2 (as in FIG. 4) can be used to construct nucleotide sequences encoding 1) the active fragments of larger proteins, or 2) fusions of multiple antimicrobial fragments. This can be done using standard codon tables and cloning methods as described in laboratory manuals such as Current Protocols in Molecular Biology (copyright 1987-1995 edited by Ausubel F. M. et al. and published by John Wiley & Sons, Inc., printed in the USA). Subsequently, these can be expressed in liquid culture for purification and testing, or the sequences can be expressed in transgenic plants after placing them in appropriate expression vectors.
- The antimicrobial proteins per se will manifest a particular three-dimensional structure which may be determined using X-ray crystallography or nuclear magnetic resonance techniques. This structure will be responsible in large part for the antimicrobial activity of the protein. The sequence of the protein can also be subjected to structure prediction algorithms to assess whether any secondary structure elements are likely to be exhibited by the protein (see Example 8 and FIG. 7). Secondary structures, thus predicted, can then be used to model three-dimensional global structures. Although three-dimensional structure prediction is not feasible for most proteins, the secondary structure predictions for MiAMP2c were sufficiently simple and clear that a three-dimensional model structure has been obtained for the MiAMP2c protein. Homologues exhibiting the same cysteine spacing and other key elements will also adopt the same three-dimensional structure. Example 8 shows that the structure most likely to be adopted by MiAMP2c (and homologues) is a helix-turn-helix structure stabilised by at least two disulfide bridges connecting the two antiparallel α-helical segments (see FIG. 8). Additional stabilisation can be provided by an extra disulfide bridge (e.g., as in MiAMP2b) or by a hydrophobic ring-stacking interaction between tyrosine and/or phenylalanine residues (e.g., MiAMP2a and MiAMP2c), each located on the same face of the α-helical segments as the normally present cysteine residues which participate in the 2 disulfide linkages mentioned above. NMR signals exhibited by MiAMP2c are consistent with the three-dimensional global model produced from the secondary-structure predictions mentioned above.
- It will be appreciated that one skilled in the art could take a protein with known structure, alter the sequence significantly, and yet retain the overall three-dimensional shape and antimicrobial activity of the protein. One aspect of the structure that most likely could not be altered without seriously affecting the structure (and, therefore, the activity of the protein) is the content and spacing of the cysteine residues since this would disrupt the formation of disulfide bonds which are critical to a) maintaining the overall structure of the protein and/or b) making the protein more resistant to denaturation and proteolysis (stabilizing the protein structure). In particular, it is essential that cysteine residues reside on one face of the helix in which they are contained. This can best be accomplished by maintaining a three-residue spacing between the cysteine residues within each helix, but, can also be accomplished with a two-residue interval between the cysteine residues—provided the cysteines on the other helical segment are separated by three residues (i.e., C-X-X-C-X-X-X-C-nX-C-X-X-X-C-X-X-X-C where C is cysteine, X is any amino acid, and n is the number of residues forming a turn between the two α-helical segments). Aromatic tyrosine (or phenylalanine) residues can also function to add stability to the protein structure if they are located on the same face of the helix as the cysteine side chains. This can be accomplished by providing appropriate spacing of two or three residues between the aromatic residue and the proximate cysteine residue (i.e., Z-X-X-C-X-X-X-C-nX-C-X-X-X-C-X-X-X-Z where Z is tyrosine or phenylalanine).
- The distribution of positive (and negative) charges on the various surfaces of the protein will also serve a critical role in determining the structure and activity of the protein. In particular, the distribution of positively-charged residues in an α-helical region of a protein can result in positive charges lying on one face of the helix or may result in the charged residues being concentrated in some particular portion of the molecule. An alternative distribution of positively charged residues is for them to project into the solvent in a radial orientation to the core of the protein. This orientation is predicted for several of the MiAMP2 homologues (data not shown). The spacing which is required for positioning of the residues on one face of the helix or the-spacing required to accomplish a radial orientation from the core can easily be determined by one skilled in the art using a helical wheel plot with the sequence of interest. A helical wheel plot uses the fact that, in α-helices, each turn of the helix is composed of 3.6 residues on average. This number translates to 100° of rotational translation per residue making it possible to construct a plot showing the distribution of side chains in a helical region. FIG. 8 shows how the spacing of charged residues can lead to most of the positively charged side chains being localised on one face of the helix. It will be appreciated by one of skill in the art that positive charges are conferred by arginine and lysine residues.
- In order for the protein to develop into a helix-turn-helix structure, it is also necessary to have particular residues that favor α-helix formation and that also favor a turn structure in the middle portion of the amino acid sequence (and disfavor a helical structure in the turn region). This can be accomplished by a proline residue or residues in the middle of the turn segment as seen with many of the MiAMP2 homologues. When proline is not present, glycine can also contribute to breaking a continuous helix structure, and inducing the formation of a turn at this position. In one case (i.e., TcAMP1), it appears that serine may be taking on this role. It will be appreciated that the residues in this region of the protein will usually favor the formation of a turn structure; residues which fulfill this requirement include proline, glycine, serine, and aspartic acid; but, other residues are also allowed.
- The DNA sequences reported here are an extremely powerful tool which can be used to obtain homologous genes from other species. Using the DNA sequences, one skilled in the art can design and synthesise oligonucleotide probes which can be used to screen cDNA libraries from other species of plants for the presence of genes encoding antimicrobial proteins homologous to the ones described here. This would simply involve construction of a cDNA library and subsequent screening of the library using as the oligonucleotide probe one or part of one of the sequences reported here (such as sequence ID. No. 2 or the PCR fragment described in Example 9). Other oligonucleotide sequences coding for proteins homologous to MiAMP2 can also be used for this purpose (e.g., DNA sequences corresponding to cotton and cocoa vicilins). Making and screening of a cDNA library can be carried out by purchasing a kit for said purpose (e.g., from Stratagene) or by following well established protocols described in available DNA cloning manuals (see Current Protocols in Molecular Biology, supra). It is relatively straight forward to construct libraries of various species and to specifically isolate vicilin homologues which are similar to the Macadamia, cotton, or cocoa vicilins by using a simple DNA hybridization technique to screen such libraries. Once cloned, these vicilin-related sequences can then be examined for the presence of MiAMP2-like subunits. Such subunits can easily be expressed in E. coli using the system described in Examples 10 and 11. Subsequently, these proteins can also be expressed in transgenic.
- Genes, or fragments thereof, under the control of a constitutive or inducible promoter, can then be cloned into a biological system which allows expression of the protein encoded thereby. Transformation methods allowing for the protein to be expressed in a variety of systems are known. The protein can thus be expressed in any suitable system for the purpose of producing the protein for further use. Suitable hosts for the expression of the protein include E. coli, fungal cells, insect cells, mammalian cells, and plants. Standard methods for expressing proteins in such hosts are described in a variety of texts including section 16 (Protein Expression) of Current Protocols in Molecular Biology (supra).
- Plant cells can be transformed with DNA constructs of the invention according to a variety of known methods (Agrobacterium, Ti plasmids, electroporation, micro-injections, micro-projectile gun, and the like). DNA sequences encoding the Macadamia integrifolia antimicrobial protein subunits (i.e. fragments a, b, c, or d from the MiAMP2 clones) as well as DNA coding for other homologues can be used in conjunction with a DNA sequence encoding a preprotein from which the mature protein is produced. This preprotein can contain a native or synthetic signal peptide sequence which will target the protein to a particular cell compartment (e.g., the apoplast or the vacuole). These coding sequences can be ligated to a plant promoter sequence that will ensure strong expression in plant cells. This promoter sequence might ensure strong constitutive expression of the protein in most or all plant cells, it may be a promoter which ensures expression in specific tissues or cells that are susceptible to microbial infection and it may also be a promoter which ensures strong induction of expression during the infection process. These types of gene cassettes will also include a transcription termination and
polyadenylation sequence 3′ of the antimicrobial protein coding region to ensure efficient production and stabilisation of the mRNA encoding the antimicrobial proteins. It is possible that efficient expression of the antimicrobial proteins disclosed herein might be facilitated by inclusion of their individual DNA sequences into a sequence encoding a much larger protein which is processed in planta to produce one or more active MiAMP2-like fragments. - Gene cassettes encoding the MiAMP2 series antimicrobial proteins (i.e., MiAMP2a, b, c, or d; or all of the subunits together; or the entire MiAMP2 clone) or homologues of the MiAMP2 proteins as described above can then be expressed in plant cells using two common methods. Firstly, the gene cassettes can be ligated into binary vectors carrying: i) left and right border sequences that flank the T-DNA of the Agrobacterium tumefaciens Ti plasmid; ii) a suitable selectable marker gene for the selection of antibiotic resistant plant cells; iii) origins of replication that function in either A. tumefaciens or Escherichia coli; and iv) antibiotic resistance genes that allow selection of plasmid-carrying cells of A. tumefaciens and E. coli. This binary vector carrying the chimaeric MiAMP2 encoding gene can be introduced by either electroporation or triparental mating into A. tumefaciens strains carrying disarmed Ti plasmids such as strains LBA4404, GV3101, and AGL1 or into A. rhizogenes strains such as A4 or NCCP1885. These Agrobacterium strains can then be co-cultivated with suitable plant explants or intact plant tissue and the transformed plant cells and/or regenerants selected using antibiotic resistance.
- A second method of gene transfer to plants can be achieved by direct insertion of the gene in target plant cells. For example, an MiAMP2-encoding gene cassette can be co-precipitated onto gold or tungsten particles along with a plasmid encoding a chimaeric gene for antibiotic resistance in plants. The tungsten particles can be accelerated using a fast flow of helium gas and the particles allowed to bombard a suitable plant tissue. This can be an embryogenic cell culture, a plant explant, a callus tissue or cell suspension or an intact meristem. Plants can be recovered using the antibiotic resistance gene for selection and antibodies used to detect plant cells expressing the MiAMP2 proteins or related fragments.
- The expression of MiAMP2 proteins in the transgenic plants can be detected using either antibodies raised to the protein(s) or using antimicrobial bioassays. These and other related methods for the expression of MiAMP2 proteins or fragments thereof in plants are described in Plant Molecular Biology (2nd ed., edited by Gelvin, S. B. and Schilperoort, R. A., ©1994, published by Kluwer Academic Publishers, Dordrecht, The Netherlands)
- Both monocotyledonous and dicotyledonous plants can be transformed and regenerated. Examples of genetically modified plants include maize, banana, peanut, field peas, sunflower, tomato, canola, tobacco, wheat, barley, oats, potato, soybeans, cotton, carnations, roses, sorghum. These, as well as other agricultural plants can be transformed with the antimicrobial genes such that they would exhibit a greater degree of resistance to pathogen attack. Alternatively, the proteins can be used for the control of diseases by topological application.
- The invention also relates to application of antimicrobial protein in the control of pathogens of mammals, including humans. The protein can be used either in topological or intravenous applications for the control of microbial infections.
- As indicated above in the description of the tenth embodiment, the invention includes within its scope the preparation of antimicrobial proteins based on the prototype MiAMP2 series of proteins. New sequences can be designed from the MiAMP2 amino acid sequences which substantially retain the distribution of positively charged residues relative to cysteine residues as found in the MiAMP2 proteins. The new sequence can be synthesised or expressed from a gene encoding the sequence in an appropriate host cell. Suitable methods for such procedures have been described above. Expression of the new protein in a genetically engineered cell will typically result in a product having a correct three-dimensional structure, including correctly formed disulphide linkages between cysteine residues. However, even if the protein is chemically synthesised, methods are known in the art for further processing of the protein to break undesireable disulfide bridges and form the bridges between the desired cysteine residues to give the desired three-dimensional structure should this be necessary.
- Macadamia integrifolia antimicrobial
proteins series number 2 - As indicated above, a new series of potent antimicrobial proteins has been identified in the seeds of Macadamia integrifolia. The proteins collectivelly are called the MiAMP2 series of antimicrobial proteins (or MiAMP2 proteins) because they are all found on one large preproprotein which is processed into smaller subunits, each exhibiting antimicrobial activity; they represent the second class of antimicrobial proteins isolated from Macadamia integrifolia. Each protein fragment of the series has a characteristic pI value. MiAMP2a, b, c, and d subunits as shown in FIG. 4 have predicted pI values of 4.4, 4.6, 11.5, and 11.6 respectively (predicted using raw sequence data without the His tag or cleavage sequences associated with expression of fragments in the vector pET16b), and contain two sets of CXXXC motifs which are important in stabilising the three-dimensional structure of the protein through the formation of disulfide bonds. Additionally, the proteins contain either an added set of aromatic (tyrosine/phenylalanine) residues or an added set of cysteine residues located at positions which would give more stability to the helix-turn-helix structure as described above and in Example 8.
- The amino acid sequences of the MiAMP2 series of proteins share significant homology with fragments of previously described proteins in sequence databases (Swiss Prot and Non-redundant databases) searched using the BLASTP algorithm (Altschul, S. F. et al. [1990] J. Mol. Biol. 215:403). In particular, MiAMP2a, b, c and d sequences exhibit significant similarity with regions of cocoa vicilin and cotton vicilin (as seen in FIG. 6). Some similarity is also seen with fragments from other seed storage proteins of peanut (Burks, A. W. et al. [1995] J. Clin. Invest. 96 (4), 1715-1721), maize (Belanger, F. C. and Kriz, A. L.[1991] Genetics 129 (3), 863-872), barley (Heck, G. R. et al. [1993] Mol. Gen. Genet. 239 (1-2), 209-218), and soybean (Sebastiani, F. L. et al. [1990] Plant Mol. Biol. 15 (1), 197-201). Although, in some cases the homology is not extremely high (for example, 18% identity between MiAMP2a and
cotton subunit 1; see FIG. 4), the spacing of the main four cysteine residues is conserved in all subunits and homologues. In addition, both cotton and cocoa vicilin-derived subunits retain the conserved tyrosine or phenylalanine residues as additional stabilisers of the tertiary structure. The cotton and cocoa vicilins with 525 and 590 amino acids, respectively, are much larger proteins than MiAMP2c (47 amino acids) (see FIGS. 4 and 6). Although MiAMP2 subunits also share some homology with MBP-1 antimicrobial protein from maize (Duvick, J. P. et al. (1992) J Biol Chem 267:18814-20) the number of residues between the CXXXC motifs is 13 which puts MBP-1 outside the specifications for the spacing given here in this application. MBP-1 is also a smaller protein (33 amino acids), overall, than the sequences claimed here and there is no evidence available the MBP-1 is derived from a larger seed storage protein other than some similarity with a portion of miaze globulin protein. However, MBP-1 cannot be derived from from the maize globulin since maize globulin contains 10 residues between the two CXXXC motifs while MBP-1 contains 13. The alignments in FIGS. 4 and 6 show the similarity in cysteine spacing between MiAMP2 subunits and the cocoa and cotton vicilin-derived molecules. The cysteine and the aromatic tyrosine/phenylalanine residues in FIGS. 4 and 6 are highlighted with bold underlined text. FIG. 4 also shows the alignment of additional proteins which can be expressed in liquid culture and shown to exhibit antimicrobial activity. - All of the MiAMP2 homologues that have been tested exhibit antifungal activity. MiAMP2 homologues show very significant inhibition of fungal growth at concentrations as low as 2 μg/ml for some of the pathogens/microbes against which the proteins were tested. Thus they can be used to provide protection against several plant diseases. MiAMP2 homologues can be used as fungicides or antibiotics by application to plant parts. The proteins can also be used to inhibit growth of pathogens by expressing them in transgenic plants. The proteins can also be used for the control of human pathogens by topological application or intravenous injection. One characteristic of the proteins is that inhibition of some microbes is suppressed by the presence of Ca 2+ (1 mM). An example of this effect is provided for MiAMP2c subunit in Table 1.
- Some of the MiAMP2 proteins and homologues could also function as insect control agents. Since some of the proteins are extremely basic (e.g., pI>11.5 for MiAMP2c and d subunits), they would maintain a strong net-positive charge even in the highly alkaline environment of an insect gut. This strong net-positive charge would enable it to interact with negatively charged structures within the gut. This interaction may lead to inefficient feeding, slowing of growth, and possibly death of the insect pest.
- Non-limiting examples of the invention follow.
- Twenty five kilograms of Mi nuts (purchased from the Macadamia Nut Factory, Queensland, Australia) were ground in a food processor (The Big Oscar, Sunbeam) and the resulting meal was extracted for 2-4 hours at 4° C. with 50 L of an ice-cold extraction buffer containing 10 mM NaH 2PO4, 15 mM Na2HPO4, 100 mM KCl, 2 mM EDTA, 0.75% polyvinylpolypyrolidone, and 0.5 mM phenylmethylsulfonyl fluoride (PMSF). The resulting homogenate was run through a kitchen strainer to remove larger particulate material and then further clarified by centrifugation (4000 rpm for 15 min) in a large capacity centrifuge. Solid ammonium sulphate was added to the supernatant to obtain 30% relative saturation and the precipitate allowed to form overnight with stirring at 4° C. Following centrifugation at 4000 rpm for 30 min, the supernatant was taken and ammonium sulphate added to achieve 70% relative saturation. The solution was allowed to precipitate overnight and then centrifuged at 4000 rpm for 30 min in order to collect the precipitated protein fraction. The precipitated protein was resuspended in a minimal volume of extraction buffer and centrifuged once again (13,000 rpm×30 min) to remove the any insoluble material yet remaining. After dialysis (10 mM ethanolamine pH 9.0, 2 mM EDTA and 1 mM PMSF) to remove residual ammonium sulphate, the protein solution was passed through a Q-Sepharose Fast Plow column (5×12 cm) previously equilibrated with 10 mM ethanolamine (pH 9), 2 mM in EDTA). The collected flowthrough from this column represents the basic (pI>9) protein fraction of the seeds. This fraction was further purified as described in Example 3.
- In general, bioassays to assess antifungal and antibacterial activity were carried out in 96-well microtitre plates. Typically, the test organism was suspended in a synthetic growth medium consisting of K 2HPO4 (2.5 mM), MgSO4 (50 μM), CaCl2 (50 μM), FeSO4 (5 μM), CoCl2 (0.1 μM), CuSO4 (0.1 μM), Na2MoO4 (2 μM), H3BO3 (0.5 μM), KI (0.1 μM), ZnSO4 (0.5 μM), MnSO4 (0.1 μM), glucose (10 g/L), asparagine (1 g/L), methionine (20 mg/L), myo-inositol (2 mg/L), biotin (0.2 mg/L), thiamine-HCl (1 mg/L) and pyridoxine-HCl (0.2 mg/L). The test organism consisted of bacterial cells, fungal spores (50,000 spores/ml) or fungal mycelial fragments (produced by blending a hyphal mass from a culture of the fungus to be tested and then filtering through a fine mesh to remove larger hyphal masses). Fifty microlitres of the test organism suspended in medium was placed into each well of the microtitre plate. A further 50 μl of the test antimicrobial solution was added to appropriate wells. To deal with well-to-well variability in the bioassay, 4 replicates of each test solution were done. Sixteen wells from each 96-well plate were used as controls for comparison with the test solutions.
- Unless otherwise stated, incubation was at 25° C. for 48 hours. All fungi including yeast were grown at 25° C. E. coli were grown at 37° C. and other bacteria were bioassayed at 28° C. Percent growth inhibition was measured by following the absorbance at 600 nm of growing cultures over various time intervals and is defined as 100 times the ratio of the average change in absorbance in the control wells minus the change in absorbance in the test well divided by the average change in absorbance at 600 nm for the control wells (i.e., [(avg change in control wells−change in test well)/(avg change in control wells)]×100). Typically, measurements were taken at 24 hour intervals and the period from 24-48 hours was used for %Inhibition measurements.
- The starting material for the isolation of the Mi antimicrobial protein was the basic fraction extracted from the mature seeds as described above in Example 1. This protein was further purified by cation exchange chromatography as shown in FIG. 1.
- About 4 g of the basic protein fraction dissolved in 20 mM sodium succinate (pH 4) was applied to an S-Sepharose High Performance column (5×60 cm) (Pharmacia) previously equilibrated with the succinate buffer. The column was eluted at 17 ml/min with a linear gradient of 20 L from 0 to 2 M NaCl in 20 mM sodium succinate (pH 4). The eluate was monitored for protein by on-line measurement of the absorbance at 280 nm and collected in 200 ml fractions. Portions of each fraction were subsequently tested in the antifungal activity assay against Phytopthora cryptogea at a concentration of 100 μg/ml in the presence and absence of 1 mM Ca2+. Results of bioassays are included in FIGS. 1a and 1 b where the elution gradient is shown as a solid line and the shaded bars represent %Inhibition. The FIG. 1a assays were conducted without added Ca2+ while 1 mM Ca2+ was included in the FIG. 1b assays. Fractionation yielded a number of unresolved peaks eluting between 0.05 and 2 M NaCl. A peak eluting at about 16 hours into the separation (fractions 90-92) showed significant antimicrobial activity.
- Fractions showing significant antimicrobial activity were further purified by reversed-phase chromatography. Aliquots of fractions 90-92 were loaded onto a Pep-S (C2/C 18), column (25×0.93 cm) (Pharmacia) equilibrated with 95% H2O/5% MeCN/0.1% TFA (=100%A). The column was eluted at 3 ml/min with a 240 ml linear gradient (80 min) from 100%A to 100%B (=5% H2O/95% MeCN/0.1% TFA). Individual peaks were collected, vacuum dried three times in order to remove traces of TFA, and subsequently resuspended in 500 microlitres of milli-Q water (Millipore Corporation water purification system) for use in bioassays as described in Example 2. FIG. 2 shows the HPLC profile of purified fraction 92 from the cation-exchange separation shown in FIGS. 1 and 2. Protein elution was monitored at 214 nm. The acetonitrile gradient is shown by the straight line. Individual peaks were bioassayed for antimicrobial activity: the bars in FIG. 3 show the inhibition corresponding to 15 μg/ml of material from each of the fractions. The active protein elutes at approximately 27 min (˜30% MeCN/0.1%TFA) and is called MiAMP2c.
- The purity of the isolated antimicrobial protein was verified by native SDS-PAGE followed by staining with coomassie blue protein staining solution. Electrophoresis was performed on a 10-20% tricine gradient gel (Novex) as per the manufacturers recommendations (100 V, 1-2 hour separation time). Under these conditions the purified MiAMP2c migrates as a single discrete band (<10 kDa in size). The detection of a single major band in the SDS-PAGE analysis together with single peaks eluting in the cation-exchange and reversed-phase separations (not shown), gives strong indication that the MiAMP2c preparation is greater than 95% pure and therefore the activity of the preparation was almost certainly due to the MiAMP2c alone and not to a minor contaminating component. A clean signal in mass spectrometric analysis (Example 5 below) also supports this conclusion.
- Purified MiAMP2c was submitted for mass spectroscopic analysis. Approximately 1 μg of protein in solution was used for testing. Analysis showed the protein to have a molecular weight of 6216.8 Da ±2 Da. Additionally, the protein was subjected to reduction of disulfide bonds with dithiothreitol and alkylation with 4-vinylpyridine. The product of this reductionlalkylation was then submitted for mass spectroscopic analysis and was shown to have gained 427 mass units (i.e. molecular weight was increased by approximately 4×10 6 Da). The gain in mass indicated that four 4-vinylpyridine groups had reacted with the reduced protein, demonstrating that the protein contains a total of 4 cysteine residues. The cysteine content has also been subsequently confirmed through amino acid sequencing.
- Approximately 1 μg of the pure protein which had been reduced and alkylated was subjected to Automated Edman degradation N-terminal sequencing. In the first sequencing run, the sequence of the first 39 residues was determined. Subsequently, approximately 1 mg of MiAMP2c was reacted with Cyanogen Bromide which cleaved the protein on the C-terminal side of Methionine-26. The C-terminal fragment generated by the cleavage reaction was purified by reversed-phase HPLC and sequenced, yielding the remaining sequence of MiAMP2c (i.e. residues 27-47). The full amino acid sequence is RQRDP QQQYE QCQER CQRHE TEPRH MQTCQ QRCER RYEKE KRKQQ KR and represents
amino acids 118 to 164 ofclone 3 from Example 9 (see FIG. 6 and SEQUENCE ID NO: 5). In the figure, cysteine residues are in bold type and underlined to facilitate recognition of the spacing patterns. Depending on the number of disulfide bonds that are formed, the protein mass will range from 6215.6 to 6219.6 Da. This is in close agreement with the mass of 6216.8 i 2 Da obtained by mass spectrometric analysis (Example 5). The measured mass closely approximates the predicted mass of MiAMP2c in a two-disulfide form as is expected to be the case. - Using standard codon tables it is possible to reverse-translate the protein sequences to obtain DNA sequences that will code for the antimicrobial proteins. The software program Mac Vector 4.5.3 was used to enter the protein sequence and obtain a degenerate nucleotide sequence. A codon usage table for tobacco was referenced in order to pick codons that would be adequately represented in tobacco for purposes of obtaining high expression in this test plant. A 30 amino-acid leader peptide was also designed to ensure efficient processing of the signal peptide and secretion of the peptide extracellularly. For this purpose, the method of Von Hiejne was used to evaluate a series of possible leader sequences for probability of cleavage at the correct position [Von Hiejne, G. (1986) Nucleic Acids Research 14(11): 4683-4690]. In particular, the amino acid sequence MAWFH VSVCN AVFVV IIIIM LLMFV PVVRG (Sequence ID. No. 11) was found to give an optimal probability of correct processing of the signal peptide immediately following the G (Gly) of this leader sequence. A 5′ untranslated region from tobacco mosaic virus was also added to this synthetic gene to promote higher translational efficiency [Dowson, M. J., et al. (1994) Plant Mol. Biol. Rep. 12(4):347-357]. The synthetic gene also contains restriction sites at the 5′ and 3′ ends and immediately 5′ of the start ATG for efficient cloning and subcloning procedures. FIG. 5 shows a synthetic DNA sequence suitable for use in plant expression experiments. In this Figure, the arrow shows where translation is initiated and the triangular symbol indicates the point of cleavage of the signal peptide.
- Using sequence analysis algorithms, putative secondary structure motifs can be assigned to the protein. Five different algorithms were used to predict whether α-helices, β-sheets, or turns can occur in the MiAMP2c protein (FIG. 4). Methods were obtained from the following sources: DPM method, Deleage, G., and Roux, B. (1987) Prot. Eng. 1:289-294; SOPMA method, Geourjon, C., and Deleage, G. (1994) Prot. Eng. 7:157-164; Gibrat method, Gibrat, J. F., Garnier, J., and Robson, B. (1987) J. Mol. Biol. 198:425-443; Levin method, Levin, J. M., Robson, B., and Garnier, J. (1986) FEBS Lett. 205:303-308; and PhD method, Rost, B., And Sander, C. (1994) Proteins 19:55-72. FIG. 7 shows the predicted locations of α-helices, β-sheets and turns. The following symbols have been used in FIG. 7: C, coil (unstructured); H, alpha helix; E, β-sheet; and S, turn. Underlined residues are those which were predicted to exhibit an α-helical structure by at least 2 separate structure prediction methods; these are represented as helices in FIG. 8.
- It is clear from the secondary structure predictions that the protein is highly α-helical. While secondary structure prediction is often difficult and inaccurate, this particular prediction gives a clear indication of the structure of the protein. Examination of the secondary-structure predictions show a clear preponderance of two α-helical regions broken by a stretch of about 5-8 residues. This is highly suggestive of a helix-turn-helix motif.
- Helical wheel analysis of the MiAMP2c amino acid sequence shows that cysteine residues with a CXXXC spacing will be aligned on one face of the helix in which they are located Since the cysteines are involved in disulfide bond formation, the cysteine side chains in one helix must form covalent bonds with the cysteine side chains located on the other helical segment. When the helical segments are arranged in such a way as to bring the cysteine side chains from each respective helix into proximity with the other cysteine side chains, the resulting three-dimensional structure is shown in FIG. 8. This structure exhibits a remarkable distribution of positively charge residues on one face of the protein comprised of two helices held together by two disulfide bonds. FIG. 8 shows how the spacing of positively charged residues in helical regions of this molecule will cause these side chains to lie on one face of the helix. The positively charged residues are the dark side chains outlined in black. Other dark side chains represent acidic residues. A proline residue (grey colour marked with a ‘P’) is located at the extreme left end of the molecule in the turn region. Solid black lines show where disulfide bonds connect the two helices. The dotted line shows where the two aromatic hydrophobic residues interact to add stability to the helix-turn-helix structure.
- This helix-turn-helix structure will be adopted by all MiAMP2 homologues containing the same cysteine spacing and residues with helix and turn-forming propensities. Other MiAMP2 fragment sequences can be superimposed onto the global structure shown in FIG. 8. The overall structure will remain essentially the same but the charge distribution will vary according to the sequences involved. In the case of MiAMP2b, the dotted line would represent an added disulfide bridge instead of a hydrophobic interaction.
- PCR Amplification of a genomic fragment of the MiAMP2c Gene
- Using the reverse-translated nucleotide sequences, degenerate primers were made for use in PCR reactions with genomic DNA from Macadamia. Primer JPM17 sequence was 5′ CAG CAG CAG TAT
GAG CAG TG 3′ and primer JPM20 degenerate sequence was 5′ TTT TTC GTA (T/T)C(T/G) (G/T)C(T/G)TTC GCA 3′ (SEQ ID NOS: 12 and 13). Primers JPM17 and JPM20 were used in PCR amplifications carried out for 30 cycles with 30 sec at 95° C., 1 min at 50° C., and 1 min at 72° C. PCR products with sizes close to those which were expected were directly sequenced (ABI PRISM Dye Terminator Cycle Sequencing Ready Reaction Kit from Perkin Elmer Corporation) after excising DNA bands from agarose gels and purifying them using a Qiagen DNA clean-up kit. Using this approach, it was possible to amplify a fragment of DNA of approximately 100 bp. Direct sequencing of this nucleotide fragment yielded the nucleotide sequence corresponding to a portion of the amino acid sequence of the antimicrobial protein MiAMP2c (amino acids 7-39 of FIG. 4). The partial nucleotide sequence obtained from the above-mentioned fragment excluding the primer sequences was 5′ TCA GAA GCG CTG CCA ACG GCG CGA GAC AGA GCC ACG ACA CAT GCA AAT TTG TCAACA ACG C 3′ (corresponding to base pairs 264 to 324 in SEQ ID NO: 6). This sequence can be used for a variety of purposes including screening of cDNA and genomic libraries for clones of MiAMP2 homologues or design of specific primers for PCR amplification reactions. - Messenger RNA Isolation from Macadamia Nut Kernels
- Fifty-eight grams of Macadamia nut kernels were ground to powder under liquid nitrogen using a mortar and pestel. RNA from ground material was then purified using a Guanidine thiocyanate/Cesium chloride technique ( Current Protocols in Molecular Biology, supra). Using this method approximately 5 mg of total RNA was isolated. Messenger RNA was then purified from total RNA using a spun column mRNA purification kit (Pharmacia).
- cDNA Library Construction
- A cDNA library was constructed in a lambda ZAP vector using a library kit from Stratagene. A total of 6 reactions were performed using 25 micrograms of messenger RNA. First and second strand cDNA synthesis was performed using MMLV Reverse transcriptase and DNA Polymerase I, respectively. After blunting the cDNA with Pfu DNA Polymerase, Eco RI linker adapters were ligated to the DNA. DNA was then kinased using T4 polynucleotide kinase and the DNA subsequently digested with Xho I restriction endonuclease. At this point cDNA material was fractionated according to size using a sephacryl-S500 column supplied with the kit. DNA was then ligated into the lambda ZAP vector. The vector containing ligated insert was then packaged into lambda phage (Gigapack III packaging extract from Stratagene).
- Screening of Library
- The library constructed above was then plated and screened in XL1-blue E. coli bacterial lawns growing in top agarose. Plaques containing individual clones were isolated by lifting onto Hybond N+ membranes (Amersham LIFE SCIENCE), hybridizing to a radiolabeled version of the genomic DNA fragment amplified above, imaging of the blot, and picking of possitive clones for the next round of screening. After secondary and tertiary screening, plaques were sufficiently isolated to allow picking of single clones. Several clones were obtained, and subsequently the pBK-CMV vector portion from the larger lambda vector was excised.
- Sequence of MiAMP2c cDNA Clones
- Vectors (pBK-CMV) containing putative MiAMP2c clones were sequenced to obtain the DNA sequence of the cloned inserts. Seven clones were partially sequenced and an additional three clones were fully sequenced (see SEQ ID NOS: 2, 4 and 6 for DNA sequences of the macadamia clones). Translation of the DNA sequences showed that the full length clones encoded highly similar proteins of 666 amino acids. FIG. 6 shows that these proteins have substantial similarity to vicilin seed-strorage proteins from cocoa and cotton. Stars show positions of conserved identities and dots show positions of conserved similarities. Examination of the protein sequences revealed that the exact MiAMP2c sequence is found within the translated protein sequence of
clone 3 atamino acid positions 118 to 164 (see FIG. 6); 1 and 2 contained sequences differing from MiAMP2c by 2 residues and 3 residues, respectively, out of 47 amino acids total in the MiAMP2c sequence.clones - The translation products of the full-length clones (i.e.,
clones 1 and 2) consist of a short signal peptide fromresidues 1 to 28, a hydrophilic region fromresidues 29 to ˜246, and then two segments stretching from residues ˜246 to 666 with a stretch of acidic residues separating them at positions 542-546. - Significantly, the hydrophilic region containing the sequence for MiAMP2c, also contains 3 additional segments which are very similar to MiAMP2 (termed MiAMP2a, b and d). These 4 segments (found between
residues 28 and ˜246) are separated by stretches in which approximately four out of five residues are acidic (usually glutamic acid). These acidic stretches occur at positions 64-68, 111-115, 171-174, and 241-246 and appear to delineate processing sites for cleavage of the 666-residue preproprotein into smaller functional fragments (acidic stretches delineating cleavage sites are shown as bold characters in FIG. 6). All four MiAMP2-like segments of the protein contain 2 doublets of cysteine residues separated by 10-12 residues to give the following pattern C-X-X-X-C-(10-12X)-C-X-X-X-C where X is any amino acid, and C is cysteine. All four segments are expected to form helix-turn-helix motifs as decribed in Example 8 above. It is clear that the cysteines in these locations will form disulfide bridges that stabilize the structure of the proteins by holding the two helical portions together. - The predicted helix-turn-helix motifs can be further stabilized in several ways. The first method of stabilization is exemplified in
segments 1 and 3 (i.e., residues 29-63 and 118-170, respectively, of the 666-residue Macadamia vicilin-like protein). These segments is the are stabilized by a hydrophobic ring-stacking interaction between two aromatic residues (one on each α-helical segment); this is normally accomplished with tyrosine residues but phenylalanine is also used. As with the cysteine residues, the location of these aromatic residues in the predicted α-helical segments is critical if they are to offer stabilization to the helix-turn-helix structure. In 1 and 3, the aromatic residues are 2 and 3 residues removed from the cysteine doublets as shown here: Z-X-X-C-X-X-X-C-(10-12X)-C-X-X-X-C-X-X-X-Z where C is cysteine and Z is usually tyrosine but can be substituted with phenylalanine as is done insegments segment 1. - The second way to stabilize the helix-turn-helix fragment is by using an added disulfide bridge as seen in fragment 2 (residues 71-110). This is accomplished by placing
2 and 3 residues removed from the cysteine doublets as shown here: nX-C-X-X-C-X-X-X-C-(10-12X)-C-X-X-X-C-X-X-X-C-nX. This is the only report that the inventors know of where a helix-turn-helix domain in an antimicrobial protein is stabilized by three disulfide bridges. While segment 4 (residues 175-241) does not contain the extra disulfide bridge or the hydrophobic ring-stacking stabilization, it is probably stabilized by means of weaker ionic and or hydrogen bonding interactions.additional cysteine residues - PCR primers flanking the nucleotide region coding for MiAMP2c were engineered to contain restriction sites for Nde I and Bam HI (corresponding to the 5′ and 3′ ends of the coding region, respectively; Primer JPM31 sequence: 5′ A CAC CAT ATG CGA CAA
CGT GAT CC 3′; Primer JPM32 sequence: 3′ C GTT GTT TTC TCT ATT CCTAGG GTT G 5′, SEQ ID NOS: 14 and 15). These primers were then used to amplify the coding region of MiAMP2c DNA. The PCR product from this amplification was then digested with Nde I and Bam HI and ligated into a pET17b vector (Novagen/Studier, F. W. et al. [1986] J Mol. Biol. 189:113) with the coding region in-frame to produce the vector pET 17-MiAMP2c. - A similar approach to the one above was used to construct vectors carrying the coding sequences of MiAMP2c homologues (i.e. MiAMP2a, b, and d as well as Tc AMP1, and TcAMP2). To construct the expression vectors for fragments a, b and d in
MiAMP2 clone 1, specific PCR primers incorporating the Nde I and Bam HI sites were designed to amplify the fragments of interest. The products were then digested with the appropriate restriction enzymes and ligated into the Nde I/Bam HI sites of a pET16b vector [Novagen] containing a His tag and a Factor Xa cleavage site (amino acid sequence MGHHH HHHHH HHSSG HIEGR HM, SEQ ID NO: 16). The protein products expressed from the pET16b vector is a fusion to the antimicrobial protein. The coding sequences for MiAMP2-like subunits from cocoa (FIG. 4, TcAMP1 and TcAMP2) were obtained from the published DNA sequence of the cocoa vicilin gene (Spencer, M. E. and Hodge R. [1992] Planta 186:567-576). Two MiAMP2-like fragments within the cocoa vicilin gene were located at the 5′ end (corresponding to the residues shown in FIG. 4), and two sets of complimentary oligonucleotides corresponding to the desired coding sequences were designed. The complimentary oligonucleotides (90 to ˜100 bases) corresponding to each cocoa subunit contained a 20 bp overlap and also contained the Nde I and Bam HI restriction endonuclease cut sites.For TcAMP, the following nucleotides were synthesised: TcAMP1 forward oligo 5′ GGGAATTCCA TATGTATGAG CGTGATCCTC GACAGCAATA CGAGCAATGC CAGAGGCGAT GCGAGTCGGA AGCGACTGAA GAAAGGGAGC 3′; TcAMP1 reverse oligo 5′ GAAGCGACTG AAGAAAGGGA GCAAGAGCAG TGTGAACAAC GCTGTGAAAG GGAGTACAAG GAGCAGCAGA GACAGCAATA GGGATCCACA C 3′. For TcAMP2, the following oligonucleotides were used: TcAMP2 forward oligo 5′ GGGAATTCCA TATGCTTCAA AGGCAATACC AGCAATGTCA AGGGCGTTGT CAAGAGCAAC AACAGGGGCA GAGAGAGCAG CAGCAGTGCC AGAGAAAATG C 3′;TcAMP2 reverse oligo 5′ GTGTGGATCC CTAGCTCCTA TTTTTTTTGT GATTATGGTA ATTCTCGTGC TCGCCTCTCT CTTGTTCCTT ATATTGCTCC CAGCATTTTC TCTGGCACTG CT 3′. - The oligonucleotide sets were added to individual PCR amplification reactions in order make individual PCR fragments containing the desired coding region. Since initial PCR amplifications gave fuzzy bands, reamplification of the original products was carried out using new 20mer primers (complimentary to the 5′ends of the forward and reverse oligonucleotides shown above) designed to amplify the entire coding region of the cocoa subunits. Once amplified, the PCR products were restriction digested with the appropriate enzymes and ligated into the vector pET16b as above. This procedure was carried out for both cocoa fragments with similarities to MiAMP2c (shown in FIG. 4).
- Starter cultures (50 ml) of E. coli strain BL21 (Grodberg, J. [1988] J. Bacteriol. 170:1245) transformed with the appropriate pET construct (Example 10) were added to 500 ml of NZCYM media (Current Protocols in Molecular Biology, supra) and cultured to an optical density of 0.6 (600 nm) and induced with the addition of 0.4 or 1.0 mM IPTG depending on whether pET17b (containing a T7 promoter) or pET16b (containing a His tag fusion and a T7 promoter/lac operator) vector was being used. After cells were induced, cultures were allowed to grow for 4 hours before harvesting. Aliquots of the growing cultures were removed at timed intervals and protein extracts run on an SDS-PAGE gel to follow the expression levels of MiAMP2 and homologues in the cultures. Fragments being expressed with a Histidine tag (i.e., in the pET16b vector), were harvested by centrifuging induced cell cultures at 5000g for 10 minutes. Cell pellets were resuspended and broken by stirring for one hour in 6 M Guanidine-HCl, buffered with 100 mM sodium phosphate and 10 mM Tris at pH 8.0. Broken cell suspensions were centrifuged at 10,000g for 20-30 minutes to settle the cellular debris. Supernatants were removed to fresh tubes and 500 mg of Ni-NTA fast flow resin (Qiagen) was added to each supternatant. After gentle mixing at 4° C. for 30-60 minutes, the suspension was loaded into a small column, rinsed two times with 8 M Urea (pH 8.0 and then pH 6.3) and subsequently, the protein was eluted using 8 M Urea pH 4.5. Protein fractions thus obtained were substantially pure but were further purified using an 9.3×250 mm C2/C18 reverse phase column (Pharmacia) and 75 minute gradient from 5% to 50% acetonitrile (0.1% TFA) flowing at 3 ml/min (data not shown).
- All of the MiAMP2c homologues (except MiAMP2c which was expressed in pET17b) were expressed in the pET16b vector containing the Histidine tag. While induction of the MiAMP2c culture proceded as above, the rest of the purification was somewhat different. In this case, MiAMP2c-expressing cells were harvested by centrifugation but were then resuspended in phosphate buffer (100 mM, pH 7.0 containing 10 mM EDTA and 1 mM PMSF) and broken open using a French press instrument. Cellular debris containing MiAMP2c inclusion bodies was solubilized using a 6 M Guanidine-HCl, 10 mM MES pH 6.0 buffer. Soluble material was then recovered after centrifugation to remove insoluble debris remaining from the solubilization step. Guanidine-HCl soluble material was then dialyzed against 10 mM MES pH 6.0 containing PMSF (1 mM) and EDTA (10 mM). Cation-exchange fractionation was carried out as described in Example 3 except on a smaller scale after the dialysis step. Subsequently, the major eluting protein from the cation-exchange column, which was MiAMP2c, was then further purified using reverse phase HPLC as described in Example 3.
- FIG. 9 shows the SDS-PAGE gel analysis of the various purification stages obtained following induction with IPTG and subsequent purification of expressed proteins. Samples analysed during the TcAMP1 purification were are as follows:
lane 1, molecular weight markers;lane 2, Ni-NTA non-binding fraction;lane 3, rinse of Ni-NTA resin withpH 8 urea;lane 4, rinse of Ni-NTA resin with pH 6.3 urea;lane 5, elution of TcAMP1 with pH 4.5 urea; andlane 6, second elution of TcAMP1 with pH 4.5 urea. TcAMP2 was purified in a similar manner and was also subjected to reverse-phase HPLC to further purify the fraction eluting from the Ni-NTA resin. FIG. 10 shows the reverse phase purification of cocoa subunit number 2 (TcAMP2). - SDS-PAGE gel analysis of the MiAMP2a, b, and d fragment purification is shown in the second panel of FIG. 9. Lane contents are as follows:
lane 1, molecular weight markers;lane 2, MiAMP2a pre-induced cellular extractp;lane 3, MiAMP2a IPTG induced cellular extract;lane 4, MiAMP2a Ni-NTA non-binding fraction;lane 5, MiAMP2a elution from Ni-NTA;lane 6, MiAMP2b pre-induced cellular extract;lane 7, MiAMP2b IPTG induced cellular extract;lane 8, MiAMP2b Ni-NTA non-binding fraction;lane 9, MiAMP2b elution from Ni-NTA;lane 10, MiAMP2d pre-induced cellular extract;lane 11, MiAMP2d IPTG induced cellular extract;lane 12, MiAMP2d Ni-NTA non-binding fraction; andlane 13, MiAMP2d elution from Ni-NTA. - Using the vectors described in Example 10, MiAMP2c, and 5 homologues (i.e., MiAMP2a, MiAMP2b, MiAMP2d, TcAMP1 and TcAMP2) were all expressed, purified and tested for antimicrobial activity. The approach taken above can be applied to all of the antimicrobial fragments described in FIG. 4. Purified fragments can then be tested for specific inhibition agains microbial pathogens of interest.
- Rabbits were immunised intramuscularly according to standard protocols with MiAMP2 conjugated to diphtheria toxoid suspended in Fruends incomplete adjuvent. Serum was harvested from the animals at regular intervals after giving the animal added doses of MiAMP2 adjuvent to boost the immune response. Approximately 100 ml of serum were collected and used for screening of crude extracts obtained from several plant seeds. One hundred gram quantities of seeds were ground and extracted to obtain a crude extract as in Example 1. Aliquots of protein were separated on SDS-PAGE gels and the gels were then blotted onto nitrocellulose membrane for subsequent detection of antibody reacting proteins. The membranes were incubated with MiAMP2c rabbit primary antibodies, washed and then incubated with alkaline phosphatase-conjugated goat anti-rabbit IgG for colorimetric detection of antigenic bands using the chemical 5-bromo-4-chloro-3-indolyl phosphate/nitroblue tetrazolium substrate system (Schleicher and Schuell). FIG. 11 shows that various other species contain immunologically-related proteins of similar size to MiAMP2c. Lanes 1-15 contain the extracts from the following species: 1) Stenocarpus sinuatus, 2) Stenocarpus sinuatus ({fraction (1/10)} loading), 3) Restio tremulus, 4) Mesomalaena tetragona, 5) Nitraria billardieri, 6) Petrophile canescens, 7) Synaphae acutiloba, 8) Dryandra formosa, 9) Lambertia inermis, 10) Stirlingia latifolia, 11) Xylomelum angustifolium, 12) Conospermum bracteosum, 13) Conospermum triplinernium, 14) Molecular weight marker, 15) Macacamia integrifolia pure MiAMP2c. Lanes 1-13 contain a variety of species, some of which show the presence of antigenically related proteins of a similar size to MiAMP2c. Other bands exhibiting higher molecular weights probably represent the larger precursor seed storage proteins from which the antimicrobial proteins are derived. Antigenically-related proteins can be seen in
1, 2, 4, 6, 7, 8, 9, and 11-13.lanes - Bioassays were also performed using crude extracts from various Proteaceae species. Specifically, extracts from Banksia robur, Banksia canei, Hakea gibbosa, Stenocarpus sinuatus, and Stirlingia latifolia have all been shown to exhibit antimicrobial activity. This activity may derive from MiAMP2 homologues since these species are related to Macadamia.
- Based on the detection of immunologically related proteins in other species of the family Proteaceae and the presence of antimicrobial activity in crude extracts, Stenocarpus sinuatis was chosen for a large scale fractionation experiment in an attempt to isolate MiAMP2c homologues. Five kg of S. sinuatus seed was frozen in liquid nitrogen and ground in a food processor (Big Oscaar Sunbeam). The ground seed was immediately placed into 12 L of 50 mM H2SO4 extraction buffer and extracted at 4° C. for 1 hour with stirring. The slurry was then centrifuged for 20 min at 10,000 g to remove particulate matter. The supernatant was then adjusted to
pH 9 using a 50 mM ammonia solution. PMSF and EDTA were added to final concentrations of 1 and 10 mM respectively. - The crude protein extract was applied to an anion exchange column (Amberlite IRA-938, Rohm and Haas) (3 cm×90 cm) equilibrated with 50 mM NH 4Ac pH 9.0 at a flow rate of 40 ml/min. The unbound protein comprising the basic protein fraction was collected and used in the subsequent purification steps.
- The basic protein fraction was adjusted to pH 5.5 with acetic acid and then applied at 10 ml/minute over 12 h to a SP-Sepharose Fast Flow (Pharmacia) Column (5 cm×60 cm) pre-equilibrated with 25 mM ammonium acetate. The column was then washed for 3.5 h with 25 mM Acetate pH 5.5. Elution of bound proteins was achieved by applying a linear gradient of NH 4Ac from 25 mM to 2.0 M (pH 5.5) at 10 ml/min over 10 h. Absorbance of the eluate was observed at 280 nm and 100 ml fractions collected (see FIG. 12).
- Cation-exchange fractions that cross-reacted with the antiserum (fractions 14-28, FIG. 12) were then further purified by reverse phase chromatography. Cross-reacting fractions were loaded onto a 7
μm Cl 8 reverse phase column (Brownlee) equilibrated with 90% H2O, 10% acetonitrile and 0.1% Trifluoroacetic acid (TFA)(=100%A). Bound proteins were eluted with a linear gradient from 100%A to 100%B (5% H2O, 95% acetonitrile, 0.08% TFA). The absorbance of the eluted proteins was monitored at 214 nm and 280 nm. The eluted proteins were dried under vacuum and resuspended in water three times to remove traces of TFA from the samples. Reverse phaseprotein elution fractions 20 to 61 were analysed by pooling 2 adjacent fractions and performing a western blot analysis (see FIG. 13). Fractions 22-41 gave a weak positive reaction and fractions 42-57 gave a strong positive reaction to the anti-MiAMP2c antiserum. Fractions that showed antifungal activity against S. sclerotiorum at 50 μg/ml and 10 μg/ml are indicated by arrows on the chromatogram. - Using the approach above, several active fractions (termed SsAMP1 and SsAMP2) were obtained which were assessed for their antifungal activity against Sclerotinia sclerotiorum, Alternaria brassicola, Leptosphaeria maculans, Verticilium dahliae and Fusarium oxysporum. Bioassays were carried out as described in Example 2 and results shown in Example 15. Another fragment which reacted with MiAMP2 antiserum was purified and sequenced (SsAMP3) but insufficient protein was available for characterisation of antimicrobial activity. Partial sequences obtained from these proteins are shown in FIG. 4 (SEQ ID NOS: 26, 27 and 28). Full sequencing of the peptides or cloning of cDNAs encoding the seed storage proteins from this species will reveal the extent of homology between these peptides and MiAMP2-series homologues.
- In an effort to determine if the full MiAMP2c molecule was absolutely necessary for the protein to exhibit antimicrobial activity, two separate peptides were chemically synthesized by Auspep Pty. Ltd. (Australia). For each peptide, the cysteine residues were changed to alanine residues so that disulfide bonds were no longer capable of being formed between two separate protein chains. Tyrosine residues were also changed to alanine since it was expected that tyrosine also participated in the helix-turn-helix stabilization and this would not be needed in the synthetic peptides lacking one of the helices. Alanine is also favorable to the formation of alpha-helices so it should not interfere with the native helical structure to a large degree. Peptide one is comprised of 22 amino acids from 118 to 139 in the amino acid sequence of clone 3 (sequence: RQRDP QQQAE QAQKR AQRRE TE, SEQUENCE ID NO: 9).
Peptide 2 is 25 amino acids in length and runs from 140 to 164 in clone 3 (sequence: PRHMQ IAQQR AERRA EKEKR KQQKR, SEQ ID NO: 10). 1 and 2 are labeled MiAMP2c pep1 and MiAMP2c pep2 respectively. These peptides were resuspended in Milli-Q water and bioassayed against a number of fungi. As seen in Table 2,Peptides peptide 2 has inhibitory activity against a variety of fungi whereaspeptide 1 exhibited little or no activity. Mixtures ofpeptide 1 andpeptide 2 exhibit similar levels of activity as seen withpeptide 2 alone indicating thatonly peptide 2 is exhibiting activity. The fact thatpeptide 2 exhibits antimicrobial activity in the absence of the helix-turn-helix structure exhibited by MiAMP2c reveals that the helix-turn-helix structure is not absolutely necessary for the peptides to retain activity. Nevertheless,peptide 2 did not exhibit the same degree of activity on a molar basis as MiAMP2c (whole fragment) indicating that the helix-turn-helix structure is important for maximal expression of antimicrobial activity by the fragments involved. It is also expected that the helix-turn-helix structure will confer greater stability to the MiAMP2 homologues, thus rendering these proteins less susceptible to proteolytic cleavage and other forms of degradation. Greater stability would lead to maintaining antimicrobial activity over a longer period of time. - MiAMP2c and each of the various MiAMP2 homologues were tested against a variety of fungi as concentrations ranging from 2 to 50 μg/ml. Table 1 shows the IC 50 value of pure MiAMP2c against various fungi and bacteria. In the table, the “>50” indicates that 50% inhibition of the fungus was not achieved at 50 μg/ml which was the highest concentration tested. The abbreviation “ND” indicates that the test was not performed or that results could not be interpreted. The antimicrobial activity of MiAMP2c was also tested in the
presence bf 1 mM Ca2+ in the test medium and the IC50 values for these tests are given in the right-hand column. As can be seen in the table, the inhibitory activity of MiAMP2c is greatly reduced (although not eliminated) in the presence of Ca2+.TABLE 1 Concentrations of MiAMP2c at which 50% inhibition of growth was observed Organism IC50 (μg/ml) IC50 + Ca2+ (μg/ml) Alternaria helianthi 5-10 ND Candida albicans >50 >50 Ceratocystis paradoxa 20-50 >50 Cercospora nicotianae 5-10 5-10 Clavibacter michiganensis 50 >50 Chalara elegans 2-5 10-20 Fusarium oxysporum 10 20-50 Sclerotinia sclerotiorum 20-50 >50 Phytophthora cryptogea 5-10 10-25 Phytophthora parasitica nicotiana 10-20 >50 Verticillium dahliae 5-10 >50 Ralstonia solanacearum >50 >50 Pseudomonas syringae tabaci >50 >50 Saccharomyces cerevisiae 20-50 >50 Escherichia coli >50 >50 - Table 2 shows the the antimicrobial activity of various homologues and fragments of MiAMP2c. In the table, the following abbreviations are used: Ab, Alternaria brassicola; Cp: Ceratocystis paradoxa; Foc: Fusarium oxysporum; Lm: Leptosphaeria maculans; Ss: Sclerotinia sclerotiorum; Vd: Verticillium dahlias. The “>50” indicates that concentrations higher than 50 μg/ml were not tested so that an IC50 value could not be established. A blank space indicates that the test was not performed or that results could not be interpreted.
- The TcAMP1 and 2 used for the results presented in Table 2 were derived from cocoa vicilin (Examples 10 and 11). SsAMP1 and 2 show reactivity with MiAMP2c antibodies and also exhibit antimicrobial activity as seen in the table below. The versions of MiAMP2a, b and d as well as TcAMP1 and TcAMP2 tested in the bioassays all contain a His tag fusion resulting from expression in the vector pET16b. MiAMP2c pep1 and 2 are the N and C terminal regions, respectively, of MiAMP2c antimicrobial peptide as specified in Example 14 above. The concentration value listed for ‘MiAMP2c pep1+2’ is the concentration of each individual peptide in the mixture. It should be remembered that MiAMP2c pep1 and pep2 are both about ½ the size of MiAMP2c; comparisons of the activity of these peptides with the MiAMP2c protein should, therefore, be made on a molar basis rather than on a strict μg/ml concentration basis. Peptides were only tested in media A which did not contain added Ca 2+.
TABLE 2 IC50 values (μg/ml) of MiAMP2 related proteins against various fungi Peptide Fungus used in bioassy tested Ab Cp Foc Lm Ss Vd MiAMP2a 5-10 2.5-5 5-10 MiAMP2b 2.5 2.5 5-10 MiAMP2c 20-50 10 20-50 5-10 MiAMP2d 5 2.5 5-10 MiAMP2c 100 >50 pep1 MiAMP2c 10-20 10-20 50 10-20 pep2 MiAMP1c 10-25 50 pep1 + 2 TcAMP1 10 5-10 2-5 10 5-20 TcAMP2 5-10 5-10 2-5 5 5-20 SsAMP1 20-50 20-50 20-50 10-20 SsAMP2 20-50 >50 >50 >50 >50 - It is worthy of note that while the TcAMP1 and 2 sequences are readily available in the public data bases, no antimicrobial activity had ever been assigned to them. These sequences were derived from much larger proteins involved in seed storage functions. The inventors have thus described a completely new activity for a small portion of the overall cocoa vicilin molecules. The activity of
1, 2, and 3 has been exemplified by other authors (Chung, R. P. T. et al. [1997] Plant Science 127:1-16).cotton fragments - The expression vector pPCV91-MiAMP2c (FIG. 14) contains the full coding region of the MiAMP2c (Example 7) DNA flanked at it 5′ end by the strong constitutive promoter of 35S RNA from the cauliflower mosaic virus (pCaMV35S) (Odel et al., [1985] Nature 313: 810-812) with a quadruple-repeat enhancer element (e-35S) to allow for high transcriptional activity (Kay et al. [1987] Science 236:1299-1302). The coding region of MiAMP2c DNA is flanked at its 3′ end by the polyadenylation sequence of 35S RNA of the cauliflower mosaic virus (pA35S). The plasmid backbone of this vector is the plasmid pPCV91 (Walden, R. et al. [1990] Methods Mol. Cell. Biol. 1:175-194). The plasmid also contains other elements useful for plant transformation such as an ampicillin resistance gene (bla) and a hygromycin resistance gene (hph) driven by the nos promoter (pnos). These and other features allow for selection in various cloning and transformation procedures. The plasmid pPCV91-MiAMP2c was constructed as follows: A cloned fragment encoding MiAMP2c (Example 7) was digested using restriction enzymes to release the MiAMP2c gene fragment containing a synthetic leader sequence. The binary vector pPCV91 was digested with the restriction enzyme Bam HI. Both the MiAMP2c DNA fragment containing and the binary vector were ligated using T4 DNA ligase to produce pPCV9 1-MiAMP2c binary vector for plant transformation (FIG. 12).
- Using this approach, other homologues of MiAMP2c can be expressed in plants. Not only can individual homologues be expressed, but they may be expressed in combination with other proteins as fusion proteins or as portions of larger precursor proteins. For example, it is possible to express the N-terminal region of MiAMP2 clone 1 (
amino acids 1 to ≈246) which contains a signal peptide and the hydrophilic region containing four antimicrobial segments. Transgenic plants can then be assessed to examine whether the individual fragments are being processed into the expected fragments by the processing machinery already present in the plant cells. It is also possible to express the entire MiAMP2 clone 1 (amino acids 1 to 666) and to examine the processing of the entire protein when expressed in transgenic plants. Homologous regions from other sequences can also be used in multiple combinations with, for example, ten (10) or more MiAMP2-like fragments expressed as one large fusion protein with acidic cleavage sites located as proper locations between each of the fragments. As well as linking MiAMP2 fragments together, it would also be possible to link MiAMP2 fragments to other useful proteins for expression in plants. - The disarmed Agrobacterium tumefaciens strain GV3101 (pMP90RK) (Koncz, Cs.[1986] Mol. Gen. Genet. 204:383-396) was transformed with the vector pPCV91-MiAMP2c (Example 16) using the method of Walkerpeach et al. (Plant Mol. Biol. Manual B1:1-19 [1994]) adapted from Van Haute et al (EMBO J. 2:411-417 1983]).
- Tobacco transformation was carried out using leaf discs of Nicotiana tabacum based on the method of Horsch et al. (Science 227:1229-1231 [1985]) and co-culturing strains containing pPCV91-MiAMP2c. After co-cultivation of Agrobacterium and tobacco leaf disks, transgenic plants (transformed with pPCV91-MiAMP2c) were regenerated on media containing 50 μg/ml hygromycin and 500 μg/ml Cefotaxime. These transgenic plants were analysed for expression of the newly-introduced genes using standard western blotting techniques (FIG. 15). FIG. 15 shows a western blot of extracts from trangenic tobacco carrying the construct for MiAMP2c from example 16.
Lane 1 contains pure MiAMP2c as a standard, 2 and 3 contain extracts from transgenic plants canying the pPCV9 1-MiAMP2c construct. As can be see in the figure, faint bands are present at approximately the correct molecular weight, indicating that the transgenic plants appear to be expressing the MiAMP2c protein. Plants capable of constitutive expression of the introduced genes may be selected and self-pollinated to give seed. Fl seedlings of the transgenic plants may be further analysed.lanes - Every homologue of MiAMP2c that has been tested has exhibited some antimicrobial activity. This evidence indicates that other homologues will also exhibit antimicrobial activity. These homologues include fragments from 1) peanut (Burks, A. W. et al. [1995] J. Clin. Invest. 96 (4), 1715-1721), 2) maize (Belanger, F. C. and Kriz, A. L.[1991] Genetics 129 (3), 863-872), 3) barley (Heck, G. R. et al. [1993] Mol. Gen. Genet. 239 (1-2), 209-218), and 4) soybean (Sebastiani, F. L. et al. [1990] Plant Mol. Biol. 15(1), 197-201). (see SEQ ID NOS: 21,22,24, and 25). Other sequences derived from seed storage proteins of the 7S class are also expected to yield homologues of MiAMP2 proteins.
-
1 40 1 666 PRT Macadamia integrifolia 1 Met Ala Ile Asn Thr Ser Asn Leu Cys Ser Leu Leu Phe Leu Leu Ser 1 5 10 15 Leu Phe Leu Leu Ser Thr Thr Val Ser Leu Ala Glu Ser Glu Phe Asp 20 25 30 Arg Gln Glu Tyr Glu Glu Cys Lys Arg Gln Cys Met Gln Leu Glu Thr 35 40 45 Ser Gly Gln Met Arg Arg Cys Val Ser Gln Cys Asp Lys Arg Phe Glu 50 55 60 Glu Asp Ile Asp Trp Ser Lys Tyr Asp Asn Gln Glu Asp Pro Gln Thr 65 70 75 80 Glu Cys Gln Gln Cys Gln Arg Arg Cys Arg Gln Gln Glu Ser Gly Pro 85 90 95 Arg Gln Gln Gln Tyr Cys Gln Arg Arg Cys Lys Glu Ile Cys Glu Glu 100 105 110 Glu Glu Glu Tyr Asn Arg Gln Arg Asp Pro Gln Gln Gln Tyr Glu Gln 115 120 125 Cys Gln Lys His Cys Gln Arg Arg Glu Thr Glu Pro Arg His Met Gln 130 135 140 Thr Cys Gln Gln Arg Cys Glu Arg Arg Tyr Glu Lys Glu Lys Arg Lys 145 150 155 160 Gln Gln Lys Arg Tyr Glu Glu Gln Gln Arg Glu Asp Glu Glu Lys Tyr 165 170 175 Glu Glu Arg Met Lys Glu Glu Asp Asn Lys Arg Asp Pro Gln Gln Arg 180 185 190 Glu Tyr Glu Asp Cys Arg Arg Arg Cys Glu Gln Gln Glu Pro Arg Gln 195 200 205 Gln His Gln Cys Gln Leu Arg Cys Arg Glu Gln Gln Arg Gln His Gly 210 215 220 Arg Gly Gly Asp Met Met Asn Pro Gln Arg Gly Gly Ser Gly Arg Tyr 225 230 235 240 Glu Glu Gly Glu Glu Glu Gln Ser Asp Asn Pro Tyr Tyr Phe Asp Glu 245 250 255 Arg Ser Leu Ser Thr Arg Phe Arg Thr Glu Glu Gly His Ile Ser Val 260 265 270 Leu Glu Asn Phe Tyr Gly Arg Ser Lys Leu Leu Arg Ala Leu Lys Asn 275 280 285 Tyr Arg Leu Val Leu Leu Glu Ala Asn Pro Asn Ala Phe Val Leu Pro 290 295 300 Thr His Leu Asp Ala Asp Ala Ile Leu Leu Val Ile Gly Gly Arg Gly 305 310 315 320 Ala Leu Lys Met Ile His His Asp Asn Arg Glu Ser Tyr Asn Leu Glu 325 330 335 Cys Gly Asp Val Ile Arg Ile Pro Ala Gly Thr Thr Phe Tyr Leu Ile 340 345 350 Asn Arg Asp Asn Asn Glu Arg Leu His Ile Ala Lys Phe Leu Gln Thr 355 360 365 Ile Ser Thr Pro Gly Gln Tyr Lys Glu Phe Phe Pro Ala Gly Gly Gln 370 375 380 Asn Pro Glu Pro Tyr Leu Ser Thr Phe Ser Lys Glu Ile Leu Glu Ala 385 390 395 400 Ala Leu Asn Thr Gln Thr Glu Lys Leu Arg Gly Val Phe Gly Gln Gln 405 410 415 Arg Glu Gly Val Ile Ile Arg Ala Ser Gln Glu Gln Ile Arg Glu Leu 420 425 430 Thr Arg Asp Asp Ser Glu Ser Arg His Trp His Ile Arg Arg Gly Gly 435 440 445 Glu Ser Ser Arg Gly Pro Tyr Asn Leu Phe Asn Lys Arg Pro Leu Tyr 450 455 460 Ser Asn Lys Tyr Gly Gln Ala Tyr Glu Val Lys Pro Glu Asp Tyr Arg 465 470 475 480 Gln Leu Gln Asp Met Asp Leu Ser Val Phe Ile Ala Asn Val Thr Gln 485 490 495 Gly Ser Met Met Gly Pro Phe Phe Asn Thr Arg Ser Thr Lys Val Val 500 505 510 Val Val Ala Ser Gly Glu Ala Asp Val Glu Met Ala Cys Pro His Leu 515 520 525 Ser Gly Arg His Gly Gly Arg Gly Gly Gly Lys Arg His Glu Glu Glu 530 535 540 Glu Asp Val His Tyr Glu Gln Val Arg Ala Arg Leu Ser Lys Arg Glu 545 550 555 560 Ala Ile Val Val Leu Ala Gly His Pro Val Val Phe Val Ser Ser Gly 565 570 575 Asn Glu Asn Leu Leu Leu Phe Ala Phe Gly Ile Asn Ala Gln Asn Asn 580 585 590 His Glu Asn Phe Leu Ala Gly Arg Glu Arg Asn Val Leu Gln Gln Ile 595 600 605 Glu Pro Gln Ala Met Glu Leu Ala Phe Ala Ala Pro Arg Lys Glu Val 610 615 620 Glu Glu Ser Phe Asn Ser Gln Asp Gln Ser Ile Phe Phe Pro Gly Pro 625 630 635 640 Arg Gln His Gln Gln Gln Ser Pro Arg Ser Thr Lys Gln Gln Gln Pro 645 650 655 Leu Val Ser Ile Leu Asp Phe Val Gly Phe 660 665 2 2171 DNA Macadamia integrifolia sig_peptide (1)...(85) mat_peptide (86)...(1999) 2 atggcgatca atacatcaaa tttatgttct cttctctttc tcctttcact cttccttctg 60 tctacgacag tgtctcttgc tgaaagtgaa tttgacaggc aggaatatga ggagtgcaaa 120 cggcaatgca tgcagttgga gacatcaggc cagatgcgtc ggtgtgtgag tcagtgcgat 180 aagagatttg aagaggatat agattggtct aagtatgata accaagagga tcctcagacg 240 gaatgccaac aatgccagag gcgatgcagg cagcaggaga gtggcccacg tcagcaacaa 300 tactgccaac gacgctgcaa ggaaatatgt gaagaagaag aagaatataa ccgacaacgt 360 gatccacagc agcaatacga gcaatgtcag aagcactgcc aacggcgcga gacagagcca 420 cgtcacatgc aaacatgtca acaacgctgc gagaggagat atgaaaagga gaaacgtaag 480 caacaaaaga gatatgaaga gcaacaacgt gaagacgaag agaaatatga agagcgaatg 540 aaggaagaag ataacaaacg cgatccacaa caaagagagt acgaagactg ccggaggcgc 600 tgcgaacaac aggagccacg tcagcagcac cagtgccagc taagatgccg agagcagcag 660 aggcaacacg gccgaggtgg cgatatgatg aaccctcaga ggggaggcag cggcagatac 720 gaggagggag aagaggagca aagcgacaac ccctactact tcgacgaacg aagcttaagt 780 acaaggttca ggaccgagga aggccacatc tcagttctgg agaacttcta tggtagatcc 840 aagcttctac gcgcactaaa aaactatcgc ttggtgctcc tcgaggctaa ccccaacgcc 900 ttcgtgctcc ctacccactt ggatgcagat gccattctct tggtcatagg agggagagga 960 gccctcaaaa tgatccacca cgacaacaga gaatcctaca acctcgagtg tggagacgta 1020 atcagaatcc cagctggaac cacattctac ttaatcaacc gagacaacaa cgagaggctc 1080 cacatagcca agttcttaca gaccatatcc actcctggcc aatacaagga attcttccca 1140 gctggaggcc aaaacccaga gccgtacctc agtaccttca gcaaagagat tctcgaggct 1200 gcgctcaaca cacaaacaga gaagctgcgt ggggtgtttg gacagcaaag ggagggagtg 1260 ataattaggg cgtcacagga gcagatcagg gagttgactc gagatgactc agagtcacga 1320 cactggcata taaggagagg tggtgaatca agcaggggac cttacaatct gttcaacaaa 1380 aggccactgt actccaacaa atacggtcaa gcctacgaag tcaaacctga ggactacagg 1440 caactccaag acatggactt atcggttttc atagccaacg tcacccaggg atccatgatg 1500 ggtcccttct tcaacactag gtctacaaag gtggtagtgg tggctagtgg agaggcagat 1560 gtggaaatgg catgccctca cttgtcggga agacacggcg gccgcggtgg aggaaaaagg 1620 catgaggagg aagaggatgt gcactatgag caggttagag cacgtttgtc gaagagagag 1680 gccattgttg ttctggcagg tcatcccgtc gtcttcgttt catccggaaa cgagaacctg 1740 ctgctttttg catttggaat caatgcccaa aacaaccacg agaacttcct cgcggggaga 1800 gagaggaacg tgctgcagca gatagagcca caggcaatgg agctagcgtt tgccgctcca 1860 aggaaagagg tagaagagtc atttaacagc caggaccagt ctatcttctt tcctgggccc 1920 aggcagcacc agcaacagtc gccccgctcc accaagcaac aacagcctct cgtctccatt 1980 ctggacttcg ttggcttcta aagttccaca aaaaagagtg tgttatgtag tataggttag 2040 tagctcctag ctcggtgtat gagagtggta agagactaag acgctaaatc cctaagtaac 2100 taacctggcg agcttgcgtg tatgcaaata aagaggaaca gctttccaac tttaaaaaaa 2160 aaaaaaaaaa a 2171 3 666 PRT Macadamia integrifolia SIGNAL (1)...(28) PEPTIDE (29)...(666) 3 Met Ala Ile Asn Thr Ser Asn Leu Cys Ser Leu Leu Phe Leu Leu Ser 1 5 10 15 Leu Phe Leu Leu Ser Thr Thr Val Ser Leu Ala Glu Ser Glu Phe Asp 20 25 30 Arg Gln Glu Tyr Glu Glu Cys Lys Arg Gln Cys Met Gln Leu Glu Thr 35 40 45 Ser Gly Gln Met Arg Arg Cys Val Ser Gln Cys Asp Lys Arg Phe Glu 50 55 60 Glu Asp Ile Asp Trp Ser Lys Tyr Asp Asn Gln Asp Asp Pro Gln Thr 65 70 75 80 Asp Cys Gln Gln Cys Gln Arg Arg Cys Arg Gln Gln Glu Ser Gly Pro 85 90 95 Arg Gln Gln Gln Tyr Cys Gln Arg Arg Cys Lys Glu Ile Cys Glu Glu 100 105 110 Glu Glu Glu Tyr Asn Arg Gln Arg Asp Pro Gln Gln Gln Tyr Glu Gln 115 120 125 Cys Gln Glu Arg Cys Gln Arg His Glu Thr Glu Pro Arg His Met Gln 130 135 140 Thr Cys Gln Gln Arg Cys Glu Arg Arg Tyr Glu Lys Glu Lys Arg Lys 145 150 155 160 Gln Gln Lys Arg Tyr Glu Glu Gln Gln Arg Glu Asp Glu Glu Lys Tyr 165 170 175 Glu Glu Arg Met Lys Glu Glu Asp Asn Lys Arg Asp Pro Gln Gln Arg 180 185 190 Glu Tyr Glu Asp Cys Arg Arg Arg Cys Glu Gln Gln Glu Pro Arg Gln 195 200 205 Gln Tyr Gln Cys Gln Arg Arg Cys Arg Glu Gln Gln Arg Gln His Gly 210 215 220 Arg Gly Gly Asp Leu Ile Asn Pro Gln Arg Gly Gly Ser Gly Arg Tyr 225 230 235 240 Glu Glu Gly Glu Glu Lys Gln Ser Asp Asn Pro Tyr Tyr Phe Asp Glu 245 250 255 Arg Ser Leu Ser Thr Arg Phe Arg Thr Glu Glu Gly His Ile Ser Val 260 265 270 Leu Glu Asn Phe Tyr Gly Arg Ser Lys Leu Leu Arg Ala Leu Lys Asn 275 280 285 Tyr Arg Leu Val Leu Leu Glu Ala Asn Pro Asn Ala Phe Val Leu Pro 290 295 300 Thr His Leu Asp Ala Asp Ala Ile Leu Leu Val Thr Gly Gly Arg Gly 305 310 315 320 Ala Leu Lys Met Ile His Arg Asp Asn Arg Glu Ser Tyr Asn Leu Glu 325 330 335 Cys Gly Asp Val Ile Arg Ile Pro Ala Gly Thr Thr Phe Tyr Leu Ile 340 345 350 Asn Arg Asp Asn Asn Glu Arg Leu His Ile Ala Lys Phe Leu Gln Thr 355 360 365 Ile Ser Thr Pro Gly Gln Tyr Lys Glu Phe Phe Pro Ala Gly Gly Gln 370 375 380 Asn Pro Glu Pro Tyr Leu Ser Thr Phe Ser Lys Glu Ile Leu Glu Ala 385 390 395 400 Ala Leu Asn Thr Gln Ala Glu Arg Leu Arg Gly Val Leu Gly Gln Gln 405 410 415 Arg Glu Gly Val Ile Ile Ser Ala Ser Gln Glu Gln Ile Arg Glu Leu 420 425 430 Thr Arg Asp Asp Ser Glu Ser Arg Arg Trp His Ile Arg Arg Gly Gly 435 440 445 Glu Ser Ser Arg Gly Pro Tyr Asn Leu Phe Asn Lys Arg Pro Leu Tyr 450 455 460 Ser Asn Lys Tyr Gly Gln Ala Tyr Glu Val Lys Pro Glu Asp Tyr Arg 465 470 475 480 Gln Leu Gln Asp Met Asp Val Ser Val Phe Ile Ala Asn Ile Thr Gln 485 490 495 Gly Ser Met Met Gly Pro Phe Phe Asn Thr Arg Ser Thr Lys Val Val 500 505 510 Val Val Ala Ser Gly Glu Ala Asp Val Glu Met Ala Cys Pro His Leu 515 520 525 Ser Gly Arg His Gly Gly Arg Arg Gly Gly Lys Arg His Glu Glu Glu 530 535 540 Glu Asp Val His Tyr Glu Gln Val Lys Ala Arg Leu Ser Lys Arg Glu 545 550 555 560 Ala Ile Val Val Pro Val Gly His Pro Val Val Phe Val Ser Ser Gly 565 570 575 Asn Glu Asn Leu Leu Leu Phe Ala Phe Gly Ile Asn Ala Gln Asn Asn 580 585 590 His Glu Asn Phe Leu Ala Gly Arg Glu Arg Asn Val Leu Gln Gln Ile 595 600 605 Glu Pro Gln Ala Met Glu Leu Ala Phe Ala Ala Pro Arg Lys Glu Val 610 615 620 Glu Glu Leu Phe Asn Ser Gln Asp Glu Ser Ile Phe Phe Pro Gly Pro 625 630 635 640 Arg Gln His Gln Gln Gln Ser Ser Arg Ser Thr Lys Gln Gln Gln Pro 645 650 655 Leu Val Ser Ile Leu Asp Phe Val Gly Phe 660 665 4 2171 DNA Macadamia integrifolia sig_peptide (1)...(86) mat_peptide (87)...(1999) 4 atggcgatca atacatcaaa tttatgttct cttctctttc tcctttccct cttccttctg 60 tcaacgacag tgtctcttgc tgaaagtgaa tttgacaggc aggaatatga ggagtgcaaa 120 cggcaatgca tgcagttgga gacatcaggc cagatgcgtc ggtgtgtgag tcagtgcgat 180 aagagatttg aagaggatat agattggtct aagtatgata accaagacga tcctcagacg 240 gattgccaac aatgccagag gcgatgcagg cagcaggaga gtggcccacg tcagcaacaa 300 tactgccaac gacgctgcaa ggaaatatgt gaagaagaag aagaatataa ccgacaacgt 360 gatccacagc agcaatacga gcaatgtcag gagcgctgcc aacggcacga gacagagcca 420 cgtcacatgc aaacatgtca acaacgctgc gagaggagat atgaaaagga gaaacgtaag 480 caacaaaaga gatatgaaga gcaacaacgt gaagacgaag agaaatatga agagcgaatg 540 aaggaagaag ataacaaacg cgatccacaa caaagagagt acgaagactg ccggaggcgc 600 tgcgaacaac aggagccacg tcagcagtac cagtgccagc gaagatgccg agagcagcag 660 aggcaacacg gccgaggtgg tgatttgatt aaccctcaga ggggaggcag cggcagatac 720 gaggagggag aagagaagca aagcgacaac ccctactact tcgacgaacg aagcttaagt 780 acaaggttca ggaccgagga aggccacatc tcagttctgg agaacttcta tggtagatcc 840 aagcttctac gcgcactaaa aaactatcgc ttggtgctcc tcgaggctaa ccccaacgcc 900 ttcgtgctcc ctacccactt ggacgcagat gccattctct tggtcaccgg agggagagga 960 gccctcaaaa tgatccaccg tgacaacaga gaatcctaca acctcgagtg tggagacgta 1020 atcagaatcc cagctggaac cacattctac ttaatcaacc gagacaacaa cgagaggctc 1080 cacatagcca agttcttaca gaccatatcc actcctggcc aatacaagga attcttccca 1140 gctggaggcc aaaacccaga gccgtacctc agtaccttca gcaaagagat tctcgaggct 1200 gcgctcaaca cacaagcaga gaggctgcgt ggggtgcttg gacagcaaag ggagggagtg 1260 ataattagtg cgtcacagga gcagatcagg gagttgactc gagatgactc agagtcacga 1320 cgctggcata taaggagagg tggtgaatca agcaggggac cttacaatct gttcaacaaa 1380 aggccactgt actccaacaa atacggtcaa gcctacgaag tcaaacctga ggactacagg 1440 caactccaag acatggacgt atcggttttc atagccaaca tcacccaggg atccatgatg 1500 ggtcccttct tcaacactag gtctacaaag gtggtagtgg tggctagtgg agaggcagat 1560 gtggaaatgg catgccctca cttgtcggga agacacggcg gccgccgtgg agggaaaagg 1620 catgaggagg aagaggatgt gcactatgag caggttaaag cacgtttgtc gaagagagag 1680 gccattgttg ttccggtagg tcatcccgtc gtcttcgttt catccggaaa cgagaacctg 1740 ctgctttttg catttggaat caatgcccaa aacaaccacg agaacttcct cgcggggaga 1800 gagaggaacg tgctgcagca gatagagcca caggcaatgg agctagcgtt tgccgctcca 1860 aggaaagagg tagaagagtt atttaacagc caggacgagt ctatcttctt tcctgggccc 1920 aggcagcacc agcaacagtc ttcccgctcc accaagcaac aacagcctct cgtctccatt 1980 ctggacttcg ttggcttcta aagttctaca aaaaagagtg tgttatgtag tataggttag 2040 tagctcctag ctcggtgtat gcgagtggta agagaccaag acgctaaatc cctaagtaac 2100 taacctggcg agcttgcgtg tatgcaaata aagaggaaca gctttccaac tttaaaaaaa 2160 aaaaaaaaaa a 2171 5 625 PRT Macadamia integrifolia PEPTIDE (1)...(625) Partial mature peptide 5 Gln Cys Met Gln Leu Glu Thr Ser Gly Gln Met Arg Arg Cys Val Ser 1 5 10 15 Gln Cys Asp Lys Arg Phe Glu Glu Asp Ile Asp Trp Ser Lys Tyr Asp 20 25 30 Asn Gln Glu Asp Pro Gln Thr Glu Cys Gln Gln Cys Gln Arg Arg Cys 35 40 45 Arg Gln Gln Glu Ser Asp Pro Arg Gln Gln Gln Tyr Cys Gln Arg Arg 50 55 60 Cys Lys Glu Ile Cys Glu Glu Glu Glu Glu Tyr Asn Arg Gln Arg Asp 65 70 75 80 Pro Gln Gln Gln Tyr Glu Gln Cys Gln Lys Arg Cys Gln Arg Arg Glu 85 90 95 Thr Glu Pro Arg His Met Gln Ile Cys Gln Gln Arg Cys Glu Arg Arg 100 105 110 Tyr Glu Lys Glu Lys Arg Lys Gln Gln Lys Arg Tyr Glu Glu Gln Gln 115 120 125 Arg Glu Asp Glu Glu Lys Tyr Glu Glu Arg Met Lys Glu Gly Asp Asn 130 135 140 Lys Arg Asp Pro Gln Gln Arg Glu Tyr Glu Asp Cys Arg Arg His Cys 145 150 155 160 Glu Gln Gln Glu Pro Arg Leu Gln Tyr Gln Cys Gln Arg Arg Cys Gln 165 170 175 Glu Gln Gln Arg Gln His Gly Arg Gly Gly Asp Leu Met Asn Pro Gln 180 185 190 Arg Gly Gly Ser Gly Arg Tyr Glu Glu Gly Glu Glu Lys Gln Ser Asp 195 200 205 Asn Pro Tyr Tyr Phe Asp Glu Arg Ser Leu Ser Thr Arg Phe Arg Thr 210 215 220 Glu Glu Gly His Ile Ser Val Leu Glu Asn Phe Tyr Gly Arg Ser Lys 225 230 235 240 Leu Leu Arg Ala Leu Lys Asn Tyr Arg Leu Val Leu Leu Glu Ala Asn 245 250 255 Pro Asn Ala Phe Val Leu Pro Thr His Leu Asp Ala Asp Ala Ile Leu 260 265 270 Leu Val Ile Gly Gly Arg Gly Ala Leu Lys Met Ile His Arg Asp Asn 275 280 285 Arg Glu Ser Tyr Asn Leu Glu Cys Gly Asp Val Ile Arg Ile Pro Ala 290 295 300 Gly Thr Thr Phe Tyr Leu Ile Asn Arg Asp Asn Asn Glu Arg Leu His 305 310 315 320 Ile Ala Lys Phe Leu Gln Thr Ile Ser Thr Pro Gly Gln Tyr Lys Glu 325 330 335 Phe Phe Pro Ala Gly Gly Gln Asn Pro Glu Pro Tyr Leu Ser Thr Phe 340 345 350 Ser Lys Glu Ile Leu Glu Ala Ala Leu Asn Thr Gln Thr Glu Arg Leu 355 360 365 Arg Gly Val Leu Gly Gln Gln Arg Glu Gly Val Ile Ile Arg Ala Ser 370 375 380 Gln Glu Gln Ile Arg Glu Leu Thr Arg Asp Asp Ser Glu Ser Arg Arg 385 390 395 400 Trp His Ile Arg Arg Gly Gly Glu Ser Ser Arg Gly Pro Tyr Asn Leu 405 410 415 Phe Asn Lys Arg Pro Leu Tyr Ser Asn Lys Tyr Gly Gln Ala Tyr Glu 420 425 430 Val Lys Pro Glu Asp Tyr Arg Gln Leu Gln Asp Met Asp Val Ser Val 435 440 445 Phe Ile Ala Asn Ile Thr Gln Gly Ser Met Met Gly Pro Phe Phe Asn 450 455 460 Thr Arg Ser Thr Lys Val Val Val Val Ala Ser Gly Glu Ala Asp Val 465 470 475 480 Glu Met Ala Cys Pro His Leu Ser Gly Arg His Gly Gly Arg Gly Gly 485 490 495 Gly Lys Arg His Glu Glu Glu Glu Glu Val His Tyr Glu Gln Val Arg 500 505 510 Ala Arg Leu Ser Lys Arg Glu Ala Ile Val Val Leu Ala Gly His Pro 515 520 525 Val Val Phe Val Ser Ser Gly Asn Glu Asn Leu Leu Leu Phe Ala Phe 530 535 540 Gly Ile Asn Ala Gln Asn Asn His Glu Asn Phe Leu Ala Gly Arg Glu 545 550 555 560 Arg Asn Val Leu Gln Gln Ile Glu Pro Gln Ala Met Glu Leu Ala Phe 565 570 575 Ala Ala Ser Arg Lys Glu Val Glu Glu Leu Phe Asn Ser Gln Asp Glu 580 585 590 Ser Ile Phe Phe Pro Gly Pro Arg Gln His Gln Gln Gln Ser Pro Arg 595 600 605 Ser Thr Lys Gln Gln Gln Pro Leu Val Ser Ile Leu Asp Phe Val Gly 610 615 620 Phe 625 6 2140 DNA Macadamia integrifolia mat_peptide (1)...(1875) partial mature peptide 6 caatgcatgc agttagagac atcaggccag atgcgtcggt gtgtgagtca gtgcgataag 60 agatttgaag aggatataga ttggtctaag tatgataacc aagaggatcc tcagacggaa 120 tgccaacaat gccagaggcg atgcaggcag caggagagtg acccacgtca gcaacaatac 180 tgccaacgac gctgcaagga aatatgtgaa gaagaagaag aatataaccg acaacgtgat 240 ccacagcagc aatacgagca atgtcagaag cgctgccaac ggcgcgagac agagccacgt 300 cacatgcaaa tatgtcaaca acgctgcgag aggagatatg aaaaggagaa acgtaagcaa 360 caaaagagat atgaagagca acaacgtgaa gacgaagaga aatatgaaga gcgaatgaag 420 gaaggagata acaaacgcga tccacaacaa agagagtacg aagactgccg gcggcactgc 480 gaacaacagg agccacgtct gcagtaccag tgccagcgaa gatgccaaga gcagcagagg 540 caacacggcc gaggtggcga tttgatgaac cctcagaggg gaggcagcgg cagatacgag 600 gagggagaag agaagcaaag cgacaacccc tactacttcg acgaacgaag cttaagtaca 660 aggttcagga ccgaggaagg ccacatctca gttctggaga acttctatgg tagatccaag 720 cttctacgcg cactaaaaaa ctatcgcttg gtgctcctcg aggctaaccc caacgccttc 780 gtgctcccta cccacttgga tgcagatgcc attctcttgg tcatcggagg gagaggagcc 840 ctcaaaatga tccaccgtga caacagagaa tcctacaacc tcgagtgtgg agacgtaatc 900 agaatcccag ctggaaccac attctactta atcaaccgag acaacaacga gaggctccac 960 atagccaagt tcttacagac catatccact cctggccaat acaaggaatt cttcccagct 1020 ggaggccaaa acccagagcc gtacctcagt accttcagca aagagattct cgaggctgcg 1080 ctcaacacac aaacagagag gctgcgtggg gtgcttggac agcaaaggga gggagtgata 1140 attagggcgt cacaggagca gatcagggag ttgactcgag atgactcaga gtcacgacgc 1200 tggcatataa ggagaggtgg tgaatcaagc aggggacctt acaatctgtt caacaaaagg 1260 ccactgtact ccaacaaata cggtcaagcc tacgaagtca aacctgagga ctacaggcaa 1320 ctccaagaca tggacgtatc agttttcata gccaacatca cccagggatc catgatgggt 1380 cccttcttca acactaggtc tacaaaggtg gtagtggtgg ctagtggaga ggcagatgtg 1440 gaaatggcat gccctcactt gtcgggaaga cacggcggcc gcggtggagg gaaaaggcat 1500 gaggaggaag aggaggtgca ctatgagcag gttagagcac gtttgtcgaa gagagaggcc 1560 attgttgttc tggcaggtca tcccgtcgtc ttcgtttcat ccggaaacga aaacctgctg 1620 ctttttgcat ttggaatcaa tgcccaaaac aaccacgaga acttcctcgc ggggagagag 1680 aggaacgtgc tgcagcagat agagccacag gcaatggagc tagcgtttgc cgcttcaagg 1740 aaagaggtag aagagttatt taacagccag gacgagtcta tcttctttcc tgggcccagg 1800 cagcaccagc aacagtcgcc ccgctccacc aagcaacaac agcctctcgt ctccattctg 1860 gacttcgttg gcttctaaag ttctacaaaa aagagtgtgt tatgtagtat aggttagtag 1920 ctcctagctc ggtgtatgag agtggtaaga gactaagacg ctaaatccct aagtaactaa 1980 cctggcgagc ttgcgtgtat gcaaataaag aggaacagct ttccaacttt agaaagctct 2040 tttttttttt ttttttcttt ctttttctta agaaataaac gaacgtagat tgcggctcaa 2100 aaaaaaaaaa aaaaaaaaaa aaaaaaaaaa aaaaaaaaaa 2140 7 525 PRT Theobroma cacao 7 Met Val Ile Ser Lys Ser Pro Phe Ile Val Leu Ile Phe Ser Leu Leu 1 5 10 15 Leu Ser Phe Ala Leu Leu Cys Ser Gly Val Ser Ala Tyr Gly Arg Lys 20 25 30 Gln Tyr Glu Arg Asp Pro Arg Gln Gln Tyr Glu Gln Cys Gln Arg Arg 35 40 45 Cys Glu Ser Glu Ala Thr Glu Glu Arg Glu Gln Glu Gln Cys Glu Gln 50 55 60 Arg Cys Glu Arg Glu Tyr Lys Glu Gln Gln Arg Gln Gln Glu Glu Glu 65 70 75 80 Leu Gln Arg Gln Tyr Gln Gln Cys Gln Gly Arg Cys Gln Glu Gln Gln 85 90 95 Gln Gly Gln Arg Glu Gln Gln Gln Cys Gln Arg Lys Cys Trp Glu Gln 100 105 110 Tyr Lys Glu Gln Glu Arg Gly Glu His Glu Asn Tyr His Asn His Lys 115 120 125 Lys Asn Arg Ser Glu Glu Glu Glu Gly Gln Gln Arg Asn Asn Pro Tyr 130 135 140 Tyr Phe Pro Lys Arg Arg Ser Phe Gln Thr Arg Phe Arg Asp Glu Glu 145 150 155 160 Gly Asn Phe Lys Ile Leu Gln Arg Phe Ala Glu Asn Ser Pro Pro Leu 165 170 175 Lys Gly Ile Asn Asp Tyr Arg Leu Ala Met Phe Glu Ala Asn Pro Asn 180 185 190 Thr Phe Ile Leu Pro His His Cys Asp Ala Glu Ala Ile Tyr Phe Val 195 200 205 Thr Asn Gly Lys Gly Thr Ile Thr Phe Val Thr His Glu Asn Lys Glu 210 215 220 Ser Tyr Asn Val Gln Arg Gly Thr Val Val Ser Val Pro Ala Gly Ser 225 230 235 240 Thr Val Tyr Val Val Ser Gln Asp Asn Gln Glu Lys Leu Thr Ile Ala 245 250 255 Val Leu Ala Leu Pro Val Asn Ser Pro Gly Lys Tyr Glu Leu Phe Phe 260 265 270 Pro Ala Gly Asn Asn Lys Pro Glu Ser Tyr Tyr Gly Ala Phe Ser Tyr 275 280 285 Glu Val Leu Glu Thr Val Phe Asn Thr Gln Arg Glu Lys Leu Glu Glu 290 295 300 Ile Leu Glu Glu Gln Arg Gly Gln Lys Arg Gln Gln Gly Gln Gln Gly 305 310 315 320 Met Phe Arg Lys Ala Lys Pro Glu Gln Ile Arg Ala Ile Ser Gln Gln 325 330 335 Ala Thr Ser Pro Arg His Arg Gly Gly Glu Arg Leu Ala Ile Asn Leu 340 345 350 Leu Ser Gln Ser Pro Val Tyr Ser Asn Gln Asn Gly Arg Phe Phe Glu 355 360 365 Ala Cys Pro Glu Asp Phe Ser Gln Phe Gln Asn Met Asp Val Ala Val 370 375 380 Ser Ala Phe Lys Leu Asn Gln Gly Ala Ile Phe Val Pro His Tyr Asn 385 390 395 400 Ser Lys Ala Thr Phe Val Val Phe Val Thr Asp Gly Tyr Gly Tyr Ala 405 410 415 Gln Met Ala Cys Pro His Leu Ser Arg Gln Ser Gln Gly Ser Gln Ser 420 425 430 Gly Arg Gln Asp Arg Arg Glu Gln Glu Glu Glu Ser Glu Glu Glu Thr 435 440 445 Phe Gly Glu Phe Gln Gln Val Lys Ala Pro Leu Ser Pro Gly Asp Val 450 455 460 Phe Val Ala Pro Ala Gly His Ala Val Thr Phe Phe Ala Ser Lys Asp 465 470 475 480 Gln Pro Leu Asn Ala Val Ala Phe Gly Leu Asn Ala Gln Asn Asn Gln 485 490 495 Arg Ile Phe Leu Ala Gly Arg Pro Phe Phe Leu Asn His Lys Gln Asn 500 505 510 Thr Asn Val Ile Lys Phe Thr Val Lys Ala Ser Ala Tyr 515 520 525 8 590 PRT Gossypium hirsutum (cotton) 8 Met Val Arg Asn Lys Ser Ala Cys Val Val Leu Leu Phe Ser Leu Phe 1 5 10 15 Leu Ser Phe Gly Leu Leu Cys Ser Ala Lys Asp Phe Pro Gly Arg Arg 20 25 30 Gly Asp Asp Asp Pro Pro Lys Arg Tyr Glu Asp Cys Arg Arg Arg Cys 35 40 45 Glu Trp Asp Thr Arg Gly Gln Lys Glu Gln Gln Gln Cys Glu Glu Ser 50 55 60 Cys Lys Ser Gln Tyr Gly Glu Lys Asp Gln Gln Gln Arg His Arg Pro 65 70 75 80 Glu Asp Pro Gln Arg Arg Tyr Glu Glu Cys Gln Gln Glu Cys Arg Gln 85 90 95 Gln Glu Glu Arg Gln Gln Pro Gln Cys Gln Gln Arg Cys Leu Lys Arg 100 105 110 Phe Glu Gln Glu Gln Gln Gln Ser Gln Arg Gln Phe Gln Glu Cys Gln 115 120 125 Gln His Cys His Gln Gln Glu Gln Arg Pro Glu Lys Lys Gln Gln Cys 130 135 140 Val Arg Glu Cys Arg Glu Lys Tyr Gln Glu Asn Pro Trp Arg Gly Glu 145 150 155 160 Arg Glu Glu Glu Ala Glu Glu Glu Glu Thr Glu Glu Gly Glu Gln Glu 165 170 175 Gln Ser His Asn Pro Phe His Phe His Arg Arg Ser Phe Gln Ser Arg 180 185 190 Phe Arg Glu Glu His Gly Asn Phe Arg Val Leu Gln Arg Phe Ala Ser 195 200 205 Arg His Pro Ile Leu Arg Gly Ile Asn Glu Phe Arg Leu Ser Ile Leu 210 215 220 Glu Ala Asn Pro Asn Thr Phe Val Leu Pro His His Cys Asp Ala Glu 225 230 235 240 Lys Ile Tyr Leu Val Thr Asn Gly Arg Gly Thr Leu Thr Phe Leu Thr 245 250 255 His Glu Asn Lys Glu Ser Tyr Asn Ile Val Pro Gly Val Val Val Lys 260 265 270 Val Pro Ala Gly Ser Thr Val Tyr Leu Ala Asn Gln Asp Asn Lys Glu 275 280 285 Lys Leu Ile Ile Ala Val Leu His Arg Pro Val Asn Asn Pro Gly Gln 290 295 300 Phe Glu Glu Phe Phe Pro Ala Gly Ser Gln Arg Pro Gln Ser Tyr Leu 305 310 315 320 Arg Ala Phe Ser Arg Glu Ile Leu Glu Pro Ala Phe Asn Thr Arg Ser 325 330 335 Glu Gln Leu Asp Glu Leu Phe Gly Gly Arg Gln Ser Arg Arg Arg Gln 340 345 350 Gln Gly Gln Gly Met Phe Arg Lys Ala Ser Gln Glu Gln Ile Arg Ala 355 360 365 Leu Ser Gln Glu Ala Thr Ser Pro Arg Glu Lys Ser Gly Glu Arg Phe 370 375 380 Ala Phe Asn Leu Leu Ser Gln Thr Pro Arg Tyr Ser Asn Gln Asn Gly 385 390 395 400 Arg Phe Phe Glu Ala Cys Pro Pro Glu Phe Arg Gln Leu Arg Asp Ile 405 410 415 Asn Val Thr Val Ser Ala Leu Gln Leu Asn Gln Gly Ser Ile Phe Val 420 425 430 Pro His Tyr Asn Ser Lys Ala Thr Phe Val Ile Leu Val Thr Glu Gly 435 440 445 Asn Gly Tyr Ala Glu Met Val Ser Pro His Leu Pro Arg Gln Ser Ser 450 455 460 Tyr Glu Glu Glu Glu Glu Glu Asp Glu Glu Glu Glu Gln Glu Gln Glu 465 470 475 480 Glu Glu Arg Arg Ser Gly Gln Tyr Arg Lys Ile Arg Ser Arg Leu Ser 485 490 495 Arg Gly Asp Ile Phe Val Val Pro Ala Asn Phe Pro Val Thr Phe Val 500 505 510 Ala Ser Gln Asn Gln Asn Leu Arg Met Thr Gly Phe Gly Leu Tyr Asn 515 520 525 Gln Asn Ile Asn Pro Asp His Asn Gln Arg Ile Phe Val Ala Gly Lys 530 535 540 Ile Asn His Val Arg Gln Trp Asp Ser Gln Ala Lys Glu Leu Ala Phe 545 550 555 560 Gly Val Ser Ser Arg Leu Val Asp Glu Ile Phe Asn Ser Asn Pro Gln 565 570 575 Glu Ser Tyr Phe Val Ser Arg Gln Arg Gln Arg Ala Ser Glu 580 585 590 9 22 PRT Artificial Sequence Peptide 1 from M. integrifolia MiAMP2c in which Cys is replaced with Ala and Tyr is replaced with Ala, MiAMP2cpep1. 9 Arg Gln Arg Asp Pro Gln Gln Gln Ala Glu Gln Ala Gln Lys Arg Ala 1 5 10 15 Gln Arg Arg Glu Thr Glu 20 10 25 PRT Artificial Sequence Peptide 2 from M. integrifolia MiAMP2c, MiAMPcpep2. 10 Pro Arg His Met Gln Ile Ala Gln Gln Arg Ala Glu Arg Arg Ala Glu 1 5 10 15 Lys Glu Lys Arg Lys Gln Gln Lys Arg 20 25 11 36 PRT Artificial Sequence Synthetic DNA sequence coding for a leader peptide. 11 Ser Glu Gln Ile Asp Asn Met Ala Trp Phe His Val Ser Val Cys Asn 1 5 10 15 Ala Val Phe Val Val Ile Ile Ile Ile Met Leu Leu Met Phe Val Pro 20 25 30 Val Val Arg Gly 35 12 20 DNA Artificial Sequence Primer JPM17 which binds to M. integrifolia MiAMP2c. 12 cagcagcagt atgagcagtg 20 13 21 DNA Artificial Sequence Primer JMP20, a degenerate primer that binds to MiAMP2-like sequences. 13 tttttcgtak ckkckttcgc a 21 14 24 DNA Artificial Sequence Primer JPM31 corresponding to the 5′ coding region of MiAMP2c and containing Nde1 and BamH1 sites. 14 acaccatatg cgacaacgtg atcc 24 15 26 DNA Artificial Sequence Primer JPM32 corresponding to the 3′ coding region of MiAMP2c and containing Nde1 and BamH1 sites. 15 cgttgttttc tctattccta gggttg 26 16 22 PRT Artificial Sequence Peptide containing His tag and Factor Xa cleavage site of PET16b vector. 16 Met Gly His His His His His His His His His His Ser Ser Gly His 1 5 10 15 Ile Glu Gly Arg His Met 20 17 90 DNA Artificial Sequence TcAMP1 forward oligonucleotide. 17 gggaattcca tatgtatgag cgtgatcctc gacagcaata cgagcaatgc cagaggcgat 60 gcgagtcgga agcgactgaa gaaagggagc 90 18 91 DNA Artificial Sequence TcAMP1 reverse oligonucleotide. 18 gaagcgactg aagaaaggga gcaagagcag tgtgaacaac gctgtgaaag ggagtacaag 60 gagcagcaga gacagcaata gggatccaca c 91 19 101 DNA Artificial Sequence TcAMP2 forward oligonucleotide. 19 gggaattcca tatgcttcaa aggcaatacc agcaatgtca agggcgttgt caagagcaac 60 aacaggggca gagagagcag cagcagtgcc agagaaaatg c 101 20 102 DNA Artificial Sequence TcAMP2 reverse oligonucleotide. 20 gtgtggatcc ctagctccta ttttttttgt gattatggta attctcgtgc tcgcctctct 60 cttgttcctt atattgctcc cagcattttc tctggcactg ct 102 21 614 PRT Peanut 21 Met Arg Gly Arg Val Ser Pro Leu Met Leu Leu Leu Gly Ile Leu Val 1 5 10 15 Leu Ala Ser Val Ser Ala Thr Gln Ala Lys Ser Pro Tyr Arg Lys Thr 20 25 30 Glu Asn Pro Cys Ala Gln Arg Cys Leu Gln Ser Cys Gln Gln Glu Pro 35 40 45 Asp Asp Leu Lys Gln Lys Ala Cys Glu Ser Arg Cys Thr Lys Leu Glu 50 55 60 Tyr Asp Pro Arg Cys Val Tyr Asp Thr Gly Ala Thr Asn Gln Arg His 65 70 75 80 Pro Pro Gly Glu Arg Thr Arg Gly Arg Gln Pro Gly Asp Tyr Asp Asp 85 90 95 Asp Arg Arg Gln Pro Arg Arg Glu Glu Gly Gly Arg Trp Gly Pro Ala 100 105 110 Glu Pro Arg Glu Arg Glu Arg Glu Glu Asp Trp Arg Gln Pro Arg Glu 115 120 125 Asp Trp Arg Arg Pro Ser His Gln Gln Pro Arg Lys Ile Arg Pro Glu 130 135 140 Gly Arg Glu Gly Glu Gln Glu Trp Gly Thr Pro Gly Ser Glu Val Arg 145 150 155 160 Glu Glu Thr Ser Arg Asn Asn Pro Phe Tyr Phe Pro Ser Arg Arg Phe 165 170 175 Ser Thr Arg Tyr Gly Asn Gln Asn Gly Arg Ile Arg Val Leu Gln Arg 180 185 190 Phe Asp Gln Arg Ser Lys Gln Phe Gln Asn Leu Gln Asn His Arg Ile 195 200 205 Val Gln Ile Glu Ala Arg Pro Asn Thr Leu Val Leu Pro Lys His Ala 210 215 220 Asp Ala Asp Asn Ile Leu Val Ile Gln Gln Gly Gln Ala Thr Val Thr 225 230 235 240 Val Ala Asn Gly Asn Asn Arg Lys Ser Phe Asn Leu Asp Glu Gly His 245 250 255 Ala Leu Arg Ile Pro Ser Gly Phe Ile Ser Tyr Ile Leu Asn Arg His 260 265 270 Asp Asn Gln Asn Leu Arg Val Ala Lys Ile Ser Met Pro Val Asn Thr 275 280 285 Pro Gly Gln Phe Glu Asp Phe Phe Pro Ala Ser Ser Arg Asp Gln Ser 290 295 300 Ser Tyr Leu Gln Gly Phe Ser Arg Asn Thr Leu Glu Ala Ala Phe Asn 305 310 315 320 Ala Glu Phe Asn Glu Ile Arg Arg Val Leu Leu Glu Glu Asn Ala Gly 325 330 335 Gly Glu Gln Glu Glu Arg Gly Gln Arg Arg Arg Ser Thr Arg Ser Ser 340 345 350 Asp Asn Glu Gly Val Ile Val Lys Val Ser Lys Glu His Val Gln Glu 355 360 365 Leu Thr Lys His Ala Lys Ser Val Ser Lys Lys Gly Ser Glu Glu Glu 370 375 380 Asp Ile Thr Asn Pro Ile Asn Leu Arg Asp Gly Glu Pro Asp Leu Ser 385 390 395 400 Asn Asn Phe Gly Arg Leu Phe Glu Val Lys Pro Asp Lys Lys Asn Pro 405 410 415 Gln Leu Gln Asp Leu Asp Met Met Leu Thr Cys Val Glu Ile Lys Glu 420 425 430 Gly Ala Leu Met Leu Pro His Phe Asn Ser Lys Ala Met Val Ile Val 435 440 445 Val Val Asn Lys Gly Thr Gly Asn Leu Glu Leu Val Ala Val Arg Lys 450 455 460 Glu Gln Gln Gln Arg Gly Arg Arg Glu Gln Glu Trp Glu Glu Glu Glu 465 470 475 480 Glu Asp Glu Glu Glu Glu Gly Ser Asn Arg Glu Val Arg Arg Tyr Thr 485 490 495 Ala Arg Leu Lys Glu Gly Asp Val Phe Ile Met Pro Ala Ala His Pro 500 505 510 Val Ala Ile Asn Ala Ser Ser Glu Leu His Leu Leu Gly Phe Gly Ile 515 520 525 Asn Ala Glu Asn Asn His Arg Ile Phe Leu Ala Gly Asp Lys Asp Asn 530 535 540 Val Ile Asp Gln Ile Glu Lys Gln Ala Lys Asp Leu Ala Phe Pro Gly 545 550 555 560 Ser Gly Glu Gln Val Glu Lys Leu Ile Lys Asn Gln Arg Glu Ser His 565 570 575 Phe Val Ser Ala Arg Pro Gln Ser Gln Ser Pro Ser Ser Pro Glu Lys 580 585 590 Glu Asp Gln Glu Glu Glu Asn Gln Gly Gly Lys Gly Pro Leu Leu Ser 595 600 605 Ile Leu Lys Ala Phe Asn 610 22 582 PRT Maize 22 Met Val Ser Ala Arg Ile Val Val Leu Leu Ala Thr Leu Leu Cys Ala 1 5 10 15 Ala Ala Ala Val Ala Ser Ser Trp Glu Asp Asp Asn His His His His 20 25 30 Gly Gly His Lys Ser Gly Gln Cys Val Arg Arg Cys Glu Asp Arg Pro 35 40 45 Trp His Gln Arg Pro Arg Cys Leu Glu Gln Cys Arg Glu Glu Glu Arg 50 55 60 Glu Lys Arg Gln Glu Arg Ser Arg His Glu Ala Asp Asp Arg Ser Gly 65 70 75 80 Glu Gly Ser Ser Glu Asp Glu Arg Glu Gln Glu Lys Glu Lys Gln Lys 85 90 95 Asp Arg Arg Pro Tyr Val Phe Asp Arg Arg Ser Phe Arg Arg Val Val 100 105 110 Arg Ser Glu Gln Gly Ser Leu Arg Val Leu Arg Pro Phe Asp Glu Val 115 120 125 Ser Arg Leu Leu Arg Gly Ile Arg Asp Tyr Arg Val Ala Val Leu Glu 130 135 140 Ala Asn Pro Arg Ser Phe Val Val Pro Ser His Thr Asp Ala His Cys 145 150 155 160 Ile Cys Tyr Val Ala Glu Gly Glu Gly Val Val Thr Thr Ile Glu Asn 165 170 175 Gly Glu Arg Arg Ser Tyr Thr Ile Lys Gln Gly His Val Phe Val Ala 180 185 190 Pro Ala Gly Ala Val Thr Tyr Leu Ala Asn Thr Asp Gly Arg Lys Lys 195 200 205 Leu Val Ile Thr Lys Ile Leu His Thr Ile Ser Val Pro Gly Glu Phe 210 215 220 Gln Phe Phe Phe Gly Pro Gly Gly Arg Asn Pro Glu Ser Phe Leu Ser 225 230 235 240 Ser Phe Ser Lys Ser Ile Gln Arg Ala Ala Tyr Lys Thr Ser Ser Asp 245 250 255 Arg Leu Glu Arg Leu Phe Gly Arg His Gly Gln Asp Lys Gly Ile Ile 260 265 270 Val Arg Ala Thr Glu Glu Gln Thr Arg Glu Leu Arg Arg His Ala Ser 275 280 285 Glu Gly Gly His Gly Pro His Trp Pro Leu Pro Pro Phe Gly Glu Ser 290 295 300 Arg Gly Pro Tyr Ser Leu Leu Asp Gln Arg Pro Ser Ile Ala Asn Gln 305 310 315 320 His Gly Gln Leu Tyr Glu Ala Asp Ala Arg Ser Phe His Asp Leu Ala 325 330 335 Glu His Asp Val Ser Val Ser Phe Ala Asn Ile Thr Ala Gly Ser Met 340 345 350 Ser Ala Pro Leu Phe Asn Thr Arg Ser Phe Lys Ile Ala Tyr Val Pro 355 360 365 Asn Gly Lys Gly Tyr Ala Glu Ile Val Cys Pro His Arg Gln Ser Gln 370 375 380 Gly Gly Glu Ser Glu Arg Glu Arg Asp Lys Gly Arg Arg Ser Glu Glu 385 390 395 400 Glu Glu Glu Glu Ser Ser Glu Glu Gln Glu Glu Ala Gly Gln Gly Tyr 405 410 415 His Thr Ile Arg Ala Arg Leu Ser Pro Gly Thr Ala Phe Val Val Pro 420 425 430 Ala Gly His Pro Phe Val Ala Val Ala Ser Arg Asp Ser Asn Leu Gln 435 440 445 Ile Val Cys Phe Glu Val His Ala Asp Arg Asn Glu Lys Val Phe Leu 450 455 460 Ala Gly Ala Asp Asn Val Leu Gln Lys Leu Asp Arg Val Ala Lys Ala 465 470 475 480 Leu Ser Phe Ala Ser Lys Ala Glu Glu Val Asp Glu Val Leu Gly Ser 485 490 495 Arg Arg Glu Lys Gly Phe Leu Pro Gly Pro Glu Glu Ser Gly Gly His 500 505 510 Glu Glu Arg Glu Gln Glu Glu Glu Glu Arg Glu Glu Arg His Gly Gly 515 520 525 Arg Gly Glu Arg Glu Arg His Gly Arg Glu Glu Arg Glu Lys Glu Glu 530 535 540 Glu Arg Glu Gly Arg His Gly Gly Arg Glu Glu Arg Glu Glu Glu Glu 545 550 555 560 Arg His Gly Arg Gly Arg Arg Glu Glu Val Ala Glu Thr Leu Met Arg 565 570 575 Met Val Thr Ala Arg Met 580 23 33 PRT Maize 23 Arg Ser Gly Arg Gly Glu Cys Arg Arg Gln Cys Leu Arg Arg His Glu 1 5 10 15 Gly Gln Pro Trp Glu Thr Gln Glu Cys Met Arg Arg Cys Arg Arg Arg 20 25 30 Gly 24 637 PRT Barley 24 Met Ala Thr Arg Ala Lys Ala Thr Ile Pro Leu Leu Phe Leu Leu Gly 1 5 10 15 Thr Ser Leu Leu Phe Ala Ala Ala Val Ser Ala Ser His Asp Asp Glu 20 25 30 Asp Asp Arg Arg Gly Gly His Ser Leu Gln Gln Cys Val Gln Arg Cys 35 40 45 Arg Gln Glu Arg Pro Arg Tyr Ser His Ala Arg Cys Val Gln Glu Cys 50 55 60 Arg Asp Asp Gln Gln Gln His Gly Arg His Glu Gln Glu Glu Glu Gln 65 70 75 80 Gly Arg Gly Arg Gly Trp His Gly Glu Gly Glu Arg Glu Glu Glu His 85 90 95 Gly Arg Gly Arg Gly Arg His Gly Glu Gly Glu Arg Glu Glu Glu His 100 105 110 Gly Arg Gly Arg Gly Arg His Gly Glu Gly Glu Arg Glu Glu Glu Arg 115 120 125 Gly Arg Gly His Gly Arg His Gly Glu Gly Glu Arg Glu Glu Glu Arg 130 135 140 Gly Arg Gly Arg Gly Arg His Gly Glu Gly Glu Arg Glu Glu Glu Glu 145 150 155 160 Gly Arg Gly Arg Gly Arg Arg Gly Glu Gly Glu Arg Asp Glu Glu Gln 165 170 175 Gly Asp Ser Arg Arg Pro Tyr Val Phe Gly Pro Arg Ser Phe Arg Arg 180 185 190 Ile Ile Gln Ser Asp His Gly Phe Val Arg Ala Leu Arg Pro Phe Asp 195 200 205 Gln Val Ser Arg Leu Leu Arg Gly Ile Arg Asp Tyr Arg Val Ala Ile 210 215 220 Met Glu Val Asn Pro Arg Ala Phe Val Val Pro Gly Phe Thr Asp Ala 225 230 235 240 Asp Gly Val Gly Tyr Val Ala Gln Gly Glu Gly Val Leu Thr Val Ile 245 250 255 Glu Asn Gly Glu Lys Arg Ser Tyr Thr Val Lys Glu Gly Asp Val Ile 260 265 270 Val Ala Pro Ala Gly Ser Ile Met His Leu Ala Asn Thr Asp Gly Arg 275 280 285 Arg Lys Leu Val Ile Ala Lys Ile Leu His Thr Ile Ser Val Pro Gly 290 295 300 Lys Phe Gln Phe Leu Ser Val Lys Pro Leu Leu Ala Ser Leu Ser Lys 305 310 315 320 Arg Val Leu Arg Ala Ala Phe Lys Thr Ser Asp Glu Arg Leu Glu Arg 325 330 335 Leu Phe Asn Gln Arg Gln Gly Gln Glu Lys Thr Arg Ser Val Ser Ile 340 345 350 Val Arg Ala Ser Glu Glu Gln Leu Arg Glu Leu Arg Arg Glu Ala Ala 355 360 365 Glu Gly Gly Gln Gly His Arg Trp Pro Leu Pro Pro Phe Arg Gly Asp 370 375 380 Ser Arg Asp Thr Phe Asn Leu Leu Glu Gln Arg Pro Lys Ile Ala Asn 385 390 395 400 Arg His Gly Arg Leu Tyr Glu Ala Asp Ala Arg Ser Phe His Ala Leu 405 410 415 Ala Asn Gln Asp Val Arg Val Ala Val Ala Asn Ile Thr Pro Gly Ser 420 425 430 Met Thr Ala Pro Tyr Leu Asn Thr Gln Ser Phe Lys Leu Ala Val Val 435 440 445 Leu Glu Gly Glu Gly Glu Val Gln Ile Val Cys Pro His Leu Gly Arg 450 455 460 Glu Ser Glu Ser Glu Arg Glu His Gly Lys Gly Arg Arg Arg Glu Glu 465 470 475 480 Glu Glu Asp Asp Gln Arg Gln Gln Arg Arg Arg Gly Ser Glu Ser Glu 485 490 495 Ser Glu Glu Glu Glu Glu Gln Gln Arg Tyr Glu Thr Val Arg Ala Arg 500 505 510 Val Ser Arg Gly Ser Ala Phe Val Val Pro Pro Gly His Pro Val Val 515 520 525 Glu Ile Ser Ser Ser Gln Gly Ser Ser Asn Leu Gln Val Val Cys Phe 530 535 540 Glu Ile Asn Ala Glu Arg Asn Glu Arg Val Trp Leu Ala Gly Arg Asn 545 550 555 560 Asn Val Ile Gly Lys Leu Gly Ser Pro Ala Gln Glu Leu Thr Phe Gly 565 570 575 Arg Pro Ala Arg Glu Val Gln Glu Val Phe Arg Ala Gln Asp Gln Asp 580 585 590 Glu Gly Phe Val Ala Gly Pro Glu Gln Gln Ser Arg Glu Gln Glu Gln 595 600 605 Glu Gln Glu Arg His Arg Arg Arg Gly Asp Arg Gly Arg Gly Asp Glu 610 615 620 Ala Val Glu Thr Phe Leu Arg Met Ala Thr Gly Ala Ile 625 630 635 25 605 PRT Soybean (Glycine max) 25 Met Met Arg Ala Arg Phe Pro Leu Leu Leu Leu Gly Leu Val Phe Leu 1 5 10 15 Ala Ser Val Ser Val Ser Phe Gly Ile Ala Tyr Trp Glu Lys Glu Asn 20 25 30 Pro Lys His Asn Lys Cys Leu Gln Ser Cys Asn Ser Glu Arg Asp Ser 35 40 45 Tyr Arg Asn Gln Ala Cys His Ala Arg Cys Asn Leu Leu Lys Val Glu 50 55 60 Lys Glu Glu Cys Glu Glu Gly Glu Ile Pro Arg Pro Arg Pro Arg Pro 65 70 75 80 Gln His Pro Glu Arg Glu Pro Gln Gln Pro Gly Glu Lys Glu Glu Asp 85 90 95 Glu Asp Glu Gln Pro Arg Pro Ile Pro Phe Pro Arg Pro Gln Pro Arg 100 105 110 Gln Glu Glu Glu His Glu Gln Arg Glu Glu Gln Glu Trp Pro Arg Lys 115 120 125 Glu Glu Lys Arg Gly Glu Lys Gly Ser Glu Glu Glu Asp Glu Asp Glu 130 135 140 Asp Glu Glu Gln Asp Glu Arg Gln Phe Pro Phe Pro Arg Pro Pro His 145 150 155 160 Gln Lys Glu Glu Arg Asn Glu Glu Glu Asp Glu Asp Glu Glu Gln Gln 165 170 175 Arg Glu Ser Glu Glu Ser Glu Asp Ser Glu Leu Arg Arg His Lys Asn 180 185 190 Lys Asn Pro Phe Leu Phe Gly Ser Asn Arg Phe Glu Thr Leu Phe Lys 195 200 205 Asn Gln Tyr Gly Arg Ile Arg Val Leu Gln Arg Phe Asn Gln Arg Ser 210 215 220 Pro Gln Leu Gln Asn Leu Arg Asp Tyr Arg Ile Leu Glu Phe Asn Ser 225 230 235 240 Lys Pro Asn Thr Leu Leu Leu Pro Asn His Ala Asp Ala Asp Tyr Leu 245 250 255 Ile Val Ile Leu Asn Gly Thr Ala Ile Leu Ser Leu Val Asn Asn Asp 260 265 270 Asp Arg Asp Ser Tyr Arg Leu Gln Ser Gly Asp Ala Leu Arg Val Pro 275 280 285 Ser Gly Thr Thr Tyr Tyr Val Val Asn Pro Asp Asn Asn Glu Asn Leu 290 295 300 Arg Leu Ile Thr Leu Ala Ile Pro Val Asn Lys Pro Gly Arg Phe Glu 305 310 315 320 Ser Phe Phe Leu Ser Ser Thr Glu Ala Gln Gln Ser Tyr Leu Gln Gly 325 330 335 Phe Ser Arg Asn Ile Leu Glu Ala Ser Tyr Asp Thr Lys Phe Glu Glu 340 345 350 Ile Asn Lys Val Leu Phe Ser Arg Glu Glu Gly Gln Gln Gln Gly Glu 355 360 365 Gln Arg Leu Gln Glu Ser Val Ile Val Glu Ile Ser Lys Glu Gln Ile 370 375 380 Arg Ala Leu Ser Lys Arg Ala Lys Ser Ser Ser Arg Lys Thr Ile Ser 385 390 395 400 Ser Glu Asp Lys Pro Phe Asn Leu Arg Ser Arg Asp Pro Ile Tyr Ser 405 410 415 Asn Lys Leu Gly Lys Phe Phe Glu Ile Thr Pro Glu Lys Asn Pro Gln 420 425 430 Leu Arg Asp Leu Asp Ile Phe Leu Ser Ile Val Asp Met Asn Glu Gly 435 440 445 Ala Leu Leu Leu Pro His Phe Asn Ser Lys Ala Ile Val Ile Leu Val 450 455 460 Ile Asn Glu Gly Asp Ala Asn Ile Glu Leu Val Gly Leu Lys Glu Gln 465 470 475 480 Gln Gln Glu Gln Gln Gln Glu Glu Gln Pro Leu Glu Val Arg Lys Tyr 485 490 495 Arg Ala Glu Leu Ser Glu Gln Asp Ile Phe Val Ile Pro Ala Gly Tyr 500 505 510 Pro Val Val Val Asn Ala Thr Ser Asn Leu Asn Phe Phe Ala Ile Gly 515 520 525 Ile Asn Ala Glu Asn Asn Gln Arg Asn Phe Leu Ala Gly Ser Gln Asp 530 535 540 Asn Val Ile Ser Gln Ile Pro Ser Gln Val Gln Glu Leu Ala Phe Pro 545 550 555 560 Gly Ser Ala Gln Ala Val Glu Lys Leu Leu Lys Asn Gln Arg Glu Ser 565 570 575 Tyr Phe Val Asp Ala Gln Pro Lys Lys Lys Glu Glu Gly Asn Lys Gly 580 585 590 Arg Lys Gly Pro Leu Ser Ser Ile Leu Arg Ala Phe Tyr 595 600 605 26 23 PRT Stenocarpus sinuatus PEPTIDE (1)...(23) Partial MiAMP2c homologous peptide. 26 Val Lys Glu Asp His Gln Phe Glu Thr Arg Gly Glu Ile Leu Glu Cys 1 5 10 15 Tyr Arg Leu Cys Gln Gln Gln 20 27 17 PRT Stenocarpus sinuatus PEPTIDE (1)...(27) Partial MiAMP2c homologous peptide. 27 Gln Lys His Arg Ser Gln Ile Leu Gly Cys Tyr Leu Xaa Cys Gln Gln Leu 1 5 10 15 28 28 PRT Stenocarpus sinuatus PEPTIDE (1)...(28) Partial MiAMP2c homologous peptide. 28 Leu Asp Pro Ile Arg Gln Gln Gln Leu Cys Gln Met Arg Cys Gln Gln 1 5 10 15 Gln Glu Lys Asp Pro Arg Gln Gln Gln Gln Cys Lys 20 25 29 368 DNA Artificial Sequence A synthetic nucleotide sequence which can be used for the expression and secretion of MiAMP2c, containing the leader sequence from SEQ ID NO11 and SEQ ID NO5. 29 aactctagag cggccgcgtc gactattttt acaacaatta ccaacaacaa caaacaacaa 60 acaacattac aattactatt tacaattaca ggatccacaa ca atg gct tgg ttc 114 Met Ala Trp Phe 1 cac gtt tct gtt tgt aac gct gtt ttc gtt gtt att att att att atg 162 His Val Ser Val Cys Asn Ala Val Phe Val Val Ile Ile Ile Ile Met 5 10 15 20 ctt ctt atg ttc gtt cct gtt gtt aga ggt aga caa aga gat cct caa 210 Leu Leu Met Phe Val Pro Val Val Arg Gly Arg Gln Arg Asp Pro Gln 25 30 35 caa caa tac gag caa tgt caa aag agg tgt caa agg aga gag act gag 258 Gln Gln Tyr Glu Gln Cys Gln Lys Arg Cys Gln Arg Arg Glu Thr Glu 40 45 50 cct aga cac atg caa att tgt cag caa agg tgt gaa agg agg tac gag 306 Pro Arg His Met Gln Ile Cys Gln Gln Arg Cys Glu Arg Arg Tyr Glu 55 60 65 aag gag aag agg aag caa caa aag agg tgaggatccg tcgacgcggc 353 Lys Glu Lys Arg Lys Gln Gln Lys Arg 70 75 cgcagatcta gacaa 368 30 77 PRT Artificial Sequence A synthetic peptide sequence which can be used for the expression and secretion of MiAMP2c containing the leader sequence from SEQ ID NO11 and peptide sequence from SEQ ID NO5. 30 Met Ala Trp Phe His Val Ser Val Cys Asn Ala Val Phe Val Val Ile 1 5 10 15 Ile Ile Ile Met Leu Leu Met Phe Val Pro Val Val Arg Gly Arg Gln 20 25 30 Arg Asp Pro Gln Gln Gln Tyr Glu Gln Cys Gln Lys Arg Cys Gln Arg 35 40 45 Arg Glu Thr Glu Pro Arg His Met Gln Ile Cys Gln Gln Arg Cys Glu 50 55 60 Arg Arg Tyr Glu Lys Glu Lys Arg Lys Gln Gln Lys Arg 65 70 75 31 27 PRT Artificial Sequence Consensus sequence for antimicrobial peptides wherein X is any amino acid. 31 Cys Xaa Xaa Cys Xaa Xaa Xaa Cys Xaa Xaa Xaa Xaa Xaa Xaa Xaa Xaa 1 5 10 15 Xaa Xaa Cys Xaa Xaa Xaa Cys XaaXaa Xaa Cys 20 2532 28 PRT Artificial Sequence Consensus sequence for antimicrobial peptides wherein X is any amino acid. 32 Cys Xaa Xaa Cys Xaa Xaa Xaa Cys Xaa Xaa Xaa Xaa Xaa Xaa Xaa Xaa 1 5 10 15 Xaa Xaa Xaa Cys Xaa Xaa Xaa Cys XaaXaa Xaa Cys 20 2533 29 PRT Artificial Sequence Consensus sequence for antimicrobial peptides wherein X is any amino acid. 33 Cys Xaa Xaa Cys Xaa Xaa Xaa Cys Xaa Xaa Xaa Xaa Xaa Xaa Xaa Xaa 1 5 10 15 Xaa Xaa Xaa Xaa Cys Xaa Xaa Xaa Cys XaaXaa Xaa Cys 20 2534 27 PRT Artificial Sequence Consensus sequence for antimicrobial peptides, wherein X is any amino acid and the first and last X are Phenylalanine or Tyrosine. 34 Xaa Xaa Xaa Cys Xaa Xaa Xaa Cys Xaa Xaa Xaa Xaa Xaa Xaa Xaa Xaa 1 5 10 15 Xaa Xaa Cys Xaa Xaa Xaa Cys XaaXaa Xaa Xaa 20 2535 28 PRT Artificial Sequence Consensus sequence for antimicrobial peptides wherein X is any amino acid and the first and last X are phenylalanine or Tyrosine. 35 Xaa Xaa Xaa Cys Xaa Xaa Xaa Cys Xaa Xaa Xaa Xaa Xaa Xaa Xaa Xaa 1 5 10 15 Xaa Xaa Xaa Cys Xaa Xaa Xaa Cys XaaXaa Xaa Xaa 20 2536 29 PRT Artificial Sequence Consensus sequence for antimicrobial peptides wherein X is any amino acid and the first and last X are phenylalanine or Tyrosine. 36 Xaa Xaa Xaa Cys Xaa Xaa Xaa Cys Xaa Xaa Xaa Xaa Xaa Xaa Xaa Xaa 1 5 10 15 Xaa Xaa Xaa Xaa Cys Xaa Xaa Xaa Cys XaaXaa Xaa Xaa 20 2537 20 PRT Artificial Sequence Consensus sequence for antimicrobial peptides wherein X is any amino acid. 37 Cys Xaa Xaa Xaa Cys Xaa Xaa Xaa Xaa Xaa Xaa Xaa Xaa Xaa Xaa Cys 1 5 10 15 XaaXaa Xaa Cys 2038 21 PRT Artificial Sequence Consensus sequence for antimicrobial peptides wherein X is any amino acid. 38 Cys Xaa Xaa Xaa Cys Xaa Xaa Xaa Xaa Xaa Xaa Xaa Xaa Xaa Xaa Xaa 1 5 10 15 Cys XaaXaa Xaa Cys 2039 22 PRT Artificial Sequence Consensus sequence for antimicrobial peptides wherein X is any amino acid. 39 Cys Xaa Xaa Xaa Cys Xaa Xaa Xaa Xaa Xaa Xaa Xaa Xaa Xaa Xaa Xaa 1 5 10 15 Xaa Cys XaaXaa Xaa Cys 2040 5 PRT Artificial Sequence Consensus sequence for antimicrobial peptides wherein X is any amino acid. 40 Cys Xaa Xaa Xaa Cys 1 5
Claims (16)
1. A protein fragment having antimicrobial activity, wherein said protein fragment is selected from:
(ii) a polypeptide having an amino acid sequence selected from:
residues 29 to 73 of SEQ ID NO: 1
residues 74 to 116 of SEQ ID NO: 1
residues 117 to 185 of SEQ ID NO: 1
residues 186 to 248 of SEQ ID NO: 1
residues 29 to 73 of SEQ ID NO: 3
residues 74 to 1 16 of SEQ ID NO: 3
residues 117 to 185 of SEQ ID NO: 3
residues 186 to 248 of SEQ ID NO: 3
residues 1 to 32 of SEQ ID NO: 5
residues 33 to 75 of SEQ ID NO: 5
residues 76 to 144 of SEQ ID NO: 5
residues 145 to 210 of SEQ ID NO: 5
residues 34 to 80 of SEQ ID NO: 7
residues 81 to 140 of SEQ ID NO: 7
residues 33 to 79 of SEQ ID NO: 8
residues 80 to 109 of SEQ ID NO: 8
residues 120 to 161 of SEQ ID NO: 8
residues 32 to 91 of SEQ ID NO: 21
residues 25 to 84 of SEQ ID NO: 22
residues 29 to 94 of SEQ ID NO: 24
residues 31 to 85 of SEQ ID NO: 25
residues 1 to 23 of SEQ ID NO: 26
residues 1 to 17 of SEQ ID NO: 27
residues 1 to 28 of SEQ ID NO: 28;
(ii) a homologue of (i);
(iii) a polypeptide containing a relative cysteine spacing of C-2X-C-3X-C-(10-12)X-C-3X-C-3X-C wherein X is any amino acid residue, and C is cysteine;
(iv) a polypeptide containing a relative cysteine and tyrosine/phenylalanine spacing of Z-2X-C-3X-C-(10-12)X-C-3X-C-3X-Z wherein X is any amino acid residue, and C is cysteine, and Z is tyrosine or phenylalanine;
(v) a polypeptide containing a relative cysteine spacing of C-3X-C-(10-12)X-C-3X-C wherein X is any amino acid residue, and C is cysteine;
(vi) a polypeptide with substantially the same spacing of positively charged residues relative to the spacing of cysteine residues as (i); and
(vii) a fragment of the polypeptide of any one of (i) to (vi) which has substantially the same antimicrobial activity as (i).
2. A protein containing at least one polypeptide fragment according to claim 1 , wherein said polypeptide fragment has a sequence selected from within a sequence comprising SEQ ID NO: 1, SEQ ID NO: 3 or SEQ ID NO: 5
3. A protein having a sequence selected from SEQ ID NO: 1, SEQ ID NO: 3 or SEQ ID NO: 5.
4. An isolated or synthetic DNA encoding a polypeptide fragment according to claim 1 .
5. The DNA according to claim 4 , wherein said DNA has a sequence selected from SEQ ID NO: 2, SEQ ID NO: 4 or SEQ ID NO: 6.
6. A DNA construct which includes a DNA according to claim 4 operatively linked to elements for the expression of said encoded protein.
7. A transgenic plant harbouring a DNA construct according to claim 6 .
8. The transgenic plant according to claim 7 , wherein said plant is a monocotyledonous plant or a dicotyledonous plant.
9. The transgenic plant according to claim 7 , wherein said plant is selected from maize, banana, peanut, field peas, sunflower, tomato, canola, tobacco, wheat, barley, oats, potato, soybeans, cotton, carnations, roses, or sorghum.
10. Reproductive material of a transgenic plant according to claim 7 .
11. A composition comprising an antimicrobial protein according to claim 1 together with an agriculturally-acceptable carrier diluent or excipient.
12. A composition comprising an antimicrobial protein according to claim 1 together with an pharmaceutically-acceptable carrier diluent or excipient.
13. A method of controlling microbial infestation of a plant, the method comprising:
i) treating said plant with an antimicrobial protein according to claim 1 or a composition according to claim 11; or
ii) introducing a DNA construct according to claim 6 into said plant.
14. A method of controlling microbial infestation of a mammalian animal, the method comprising treating the animal with an antimicrobial protein according to claim 1 or a composition according to claim 12 .
15. The method of claim 14 , wherein said mammalian animal is a human.
16. A method of preparing an antimicrobial protein, which method comprises the steps of:
a) obtaining or designing an amino acid sequence which forms a helix-turn-helix structure;
b) replacing individual residues to achieve substantially the same distribution of positively charged residues and cysteine residues as in one or more of the amino acid sequences shown in FIG. 4;
c) synthesising a protein comprising said amino acid sequence chemically or by recombinant DNA techniques in liquid culture; and
d) if necessary, forming disulphide linkages between said cysteine residues.
Priority Applications (1)
| Application Number | Priority Date | Filing Date | Title |
|---|---|---|---|
| US10/147,095 US20030171274A1 (en) | 1996-12-20 | 2002-05-15 | Antimicrobial proteins |
Applications Claiming Priority (4)
| Application Number | Priority Date | Filing Date | Title |
|---|---|---|---|
| AUPO4275 | 1996-12-20 | ||
| AUPO4275A AUPO427596A0 (en) | 1996-12-20 | 1996-12-20 | Anti-microbial protein |
| US09/331,631 US7067624B2 (en) | 1996-12-20 | 1997-12-22 | Antimicrobial proteins |
| US10/147,095 US20030171274A1 (en) | 1996-12-20 | 2002-05-15 | Antimicrobial proteins |
Related Parent Applications (2)
| Application Number | Title | Priority Date | Filing Date |
|---|---|---|---|
| PCT/AU1997/000874 Division WO1998027805A1 (en) | 1996-12-20 | 1997-12-22 | Antimicrobial proteins |
| US09/331,631 Division US7067624B2 (en) | 1996-12-20 | 1997-12-22 | Antimicrobial proteins |
Publications (1)
| Publication Number | Publication Date |
|---|---|
| US20030171274A1 true US20030171274A1 (en) | 2003-09-11 |
Family
ID=3798583
Family Applications (2)
| Application Number | Title | Priority Date | Filing Date |
|---|---|---|---|
| US09/331,631 Expired - Fee Related US7067624B2 (en) | 1996-12-20 | 1997-12-22 | Antimicrobial proteins |
| US10/147,095 Abandoned US20030171274A1 (en) | 1996-12-20 | 2002-05-15 | Antimicrobial proteins |
Family Applications Before (1)
| Application Number | Title | Priority Date | Filing Date |
|---|---|---|---|
| US09/331,631 Expired - Fee Related US7067624B2 (en) | 1996-12-20 | 1997-12-22 | Antimicrobial proteins |
Country Status (13)
| Country | Link |
|---|---|
| US (2) | US7067624B2 (en) |
| EP (1) | EP1006785B1 (en) |
| JP (1) | JP2001510995A (en) |
| KR (1) | KR20000057699A (en) |
| CN (1) | CN1244769A (en) |
| AT (1) | ATE343927T1 (en) |
| AU (1) | AUPO427596A0 (en) |
| BR (1) | BR9713772A (en) |
| CA (1) | CA2274730A1 (en) |
| DE (1) | DE69736904T2 (en) |
| ES (1) | ES2277363T3 (en) |
| NZ (1) | NZ336337A (en) |
| WO (1) | WO1998027805A1 (en) |
Families Citing this family (22)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| EP1253200A1 (en) * | 2001-04-25 | 2002-10-30 | Société des Produits Nestlé S.A. | Cocoa polypeptides and their use in the production of cocoa and chocolate flavour |
| FR2843125B1 (en) * | 2002-08-02 | 2012-11-16 | Coletica | ACTIVE PRINCIPLES STIMULATING HUMAN BETA-DEFENSIVE TYPE 2 AND / OR TYPE 3, AND COSMETIC OR PHARMACEUTICAL COMPOSITIONS COMPRISING SUCH ACTIVE INGREDIENTS |
| US7589176B2 (en) | 2006-01-25 | 2009-09-15 | Pioneer Hi-Bred International, Inc. | Antifungal polypeptides |
| DE602007012343D1 (en) | 2006-05-16 | 2011-03-17 | Du Pont | ANTIMYCOTIC POLYPEPTIDE |
| CA2653404C (en) * | 2006-05-25 | 2015-06-30 | Hexima Limited | Multi-gene expression vehicle |
| US7893199B2 (en) * | 2006-09-29 | 2011-02-22 | National University Corporation Gunma University | Peptides having neutrophil-stimulating activity |
| AU2008241364B2 (en) * | 2007-04-20 | 2013-03-21 | Hexima Limited | Modified plant defensin |
| JP2013515703A (en) * | 2009-12-23 | 2013-05-09 | バイオポリス エセ.エレ. | Production of biologically active products from cocoa having PEP enzyme inhibitory activity, and antioxidant and / or anti-neurodegenerative activity |
| EP2389799A1 (en) * | 2010-05-25 | 2011-11-30 | BioMass Booster, S.L. | Method for increasing plant biomass |
| CN110227145B (en) * | 2010-10-12 | 2024-04-02 | 肯苏墨艾姆维德-生物技术达思植物有限公司 | antibacterial protein |
| US9060972B2 (en) * | 2010-10-30 | 2015-06-23 | George Dacai Liu | Recombinant hemagglutinin protein of influenza virus and vaccine containing the same |
| MX349741B (en) | 2011-02-07 | 2017-08-10 | Hexima Ltd * | Modified plant defensins useful as anti-pathogenic agents. |
| US9060973B2 (en) * | 2011-10-22 | 2015-06-23 | George Dacai Liu | Vaccine for enveloped viruses |
| CN106995491B (en) * | 2016-01-25 | 2021-12-07 | 欧蒙医学实验诊断股份公司 | Macadamia nut allergen |
| CN108504672B (en) * | 2018-03-30 | 2022-09-27 | 南京农业大学 | Ralstonia solanacearum N477 extracellular protein PHD and coding gene and application thereof |
| CN108728508B (en) * | 2018-06-14 | 2020-11-24 | 云南省热带作物科学研究所 | A kind of preparation method of macadamia nut polypeptide with antibacterial activity |
| CN112159821B (en) * | 2020-10-09 | 2022-08-05 | 西南大学 | Application of corn elicitor peptide gene ZmPep1 in improving verticillium wilt resistance of plants |
| JP7693173B2 (en) * | 2021-09-24 | 2025-06-17 | 学校法人杏林学園 | Method for detecting macadamia nut allergen-specific IgE, in vitro diagnostic agent for diagnosing macadamia nut allergy, kit for detecting macadamia nut allergen-specific IgE, and method for detecting macadamia nut allergen |
| PE20251840A1 (en) | 2022-05-10 | 2025-07-17 | Cabosse Naturals Nv | PEPTIDE WITH ANTIMICROBIAL ACTIVITY |
| CN119301145A (en) | 2022-05-10 | 2025-01-10 | 卡波塞自然有限公司 | Peptides with antimicrobial activity |
| WO2025088113A1 (en) * | 2023-10-26 | 2025-05-01 | Cabosse Naturals Nv | Peptide for treatment of fungal infections in cocoa |
| CN118546209B (en) * | 2024-07-24 | 2025-01-21 | 东北农业大学 | An oligopeptide with DPP-IV inhibitory activity and its preparation method and application |
Citations (2)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| US5422265A (en) * | 1990-12-07 | 1995-06-06 | State Of Oregon, Acting By And Through The State Board Of Higher Education On Behalf Of The Oregon Health Sciences University | DNA sequence for the human dopamine receptor D4 and expression thereof in mammalian cells |
| US5468615A (en) * | 1993-07-01 | 1995-11-21 | The Upjohn Company | Binding assay employing a synthetic gene for D4 dopamine receptors |
Family Cites Families (5)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| FR2525592B1 (en) * | 1982-04-26 | 1984-09-14 | Pasteur Institut | |
| CA2042448A1 (en) * | 1990-06-05 | 1991-12-06 | Jonathan P. Duvick | Antimicrobial peptides and plant disease resistance based thereon |
| GB9013016D0 (en) * | 1990-06-11 | 1990-08-01 | Mars Uk Ltd | Compounds |
| HU220115B (en) * | 1991-05-24 | 2001-11-28 | Universidad Politecnica De Madrid | Antipathogenic peptides, compositions containing same, method for plant protection, and sequences coding for the peptides |
| EP0789764B1 (en) * | 1994-10-28 | 2002-01-09 | Vitaleech Bioscience N.V. | A novel family of protease inhibitors, and other biologic active substances |
-
1996
- 1996-12-20 AU AUPO4275A patent/AUPO427596A0/en not_active Abandoned
-
1997
- 1997-12-22 ES ES97948648T patent/ES2277363T3/en not_active Expired - Lifetime
- 1997-12-22 NZ NZ33633797A patent/NZ336337A/en unknown
- 1997-12-22 KR KR1019990705561A patent/KR20000057699A/en not_active Withdrawn
- 1997-12-22 BR BR9713772A patent/BR9713772A/en not_active IP Right Cessation
- 1997-12-22 CN CN97181474A patent/CN1244769A/en active Pending
- 1997-12-22 JP JP52815198A patent/JP2001510995A/en active Pending
- 1997-12-22 DE DE1997636904 patent/DE69736904T2/en not_active Expired - Fee Related
- 1997-12-22 AT AT97948648T patent/ATE343927T1/en not_active IP Right Cessation
- 1997-12-22 EP EP97948648A patent/EP1006785B1/en not_active Expired - Lifetime
- 1997-12-22 CA CA 2274730 patent/CA2274730A1/en not_active Abandoned
- 1997-12-22 WO PCT/AU1997/000874 patent/WO1998027805A1/en not_active Ceased
- 1997-12-22 US US09/331,631 patent/US7067624B2/en not_active Expired - Fee Related
-
2002
- 2002-05-15 US US10/147,095 patent/US20030171274A1/en not_active Abandoned
Patent Citations (2)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| US5422265A (en) * | 1990-12-07 | 1995-06-06 | State Of Oregon, Acting By And Through The State Board Of Higher Education On Behalf Of The Oregon Health Sciences University | DNA sequence for the human dopamine receptor D4 and expression thereof in mammalian cells |
| US5468615A (en) * | 1993-07-01 | 1995-11-21 | The Upjohn Company | Binding assay employing a synthetic gene for D4 dopamine receptors |
Also Published As
| Publication number | Publication date |
|---|---|
| WO1998027805A1 (en) | 1998-07-02 |
| JP2001510995A (en) | 2001-08-07 |
| CN1244769A (en) | 2000-02-16 |
| US20020168392A1 (en) | 2002-11-14 |
| ES2277363T3 (en) | 2007-07-01 |
| CA2274730A1 (en) | 1998-07-02 |
| KR20000057699A (en) | 2000-09-25 |
| ATE343927T1 (en) | 2006-11-15 |
| BR9713772A (en) | 2000-03-21 |
| NZ336337A (en) | 2000-05-26 |
| EP1006785A1 (en) | 2000-06-14 |
| EP1006785B1 (en) | 2006-11-02 |
| US7067624B2 (en) | 2006-06-27 |
| DE69736904D1 (en) | 2006-12-14 |
| DE69736904T2 (en) | 2007-06-21 |
| AUPO427596A0 (en) | 1997-01-23 |
| EP1006785A4 (en) | 2004-05-19 |
Similar Documents
| Publication | Publication Date | Title |
|---|---|---|
| US7067624B2 (en) | Antimicrobial proteins | |
| KR100890839B1 (en) | New proteins, genes encoding them and methods of using them | |
| JPH05294995A6 (en) | Antimicrobial inverted peptides, antimicrobial oligopeptides and other antimicrobial compositions, processes for their preparation and their use | |
| JPH05294995A (en) | Reverse antimicrobial peptide, antimicrobial oligopeptide and other antimicrobial compositions, their production and use thereof | |
| JP2000502254A (en) | Antifungal protein | |
| CA2048910C (en) | Antimicrobial peptides active against plant pathogens, their method of use and various screening methods pertaining thereto | |
| US20130219532A1 (en) | Peptides with antifungal activities | |
| HU220115B (en) | Antipathogenic peptides, compositions containing same, method for plant protection, and sequences coding for the peptides | |
| US20080032924A1 (en) | Antifungal Peptides | |
| JPH07502976A (en) | biocidal protein | |
| JP2002530274A (en) | Highly stable peptides for protease degradation | |
| AU723474B2 (en) | Antimicrobial proteins | |
| CA2378432A1 (en) | Proteins and peptides | |
| CZ289646B6 (en) | Antimicrobial proteins, recombinant DNAs encoding thereof, antimicrobial preparation containing thereof and method of fighting fungi or bacteria | |
| EP0877756B1 (en) | Anti-microbial protein | |
| US6909032B2 (en) | DNA encoding a macadamia integrifolia anti-microbial protein, constructs comprising the same and plant material comprising the constructs | |
| Woytowich et al. | Plant antifungal peptides and their use in transgenic food crops | |
| AU713909B2 (en) | Anti-microbial protein | |
| US20040087771A1 (en) | Antimicrobial peptides of the family of defensins, polynucleotides encoding said peptides, transformed vectors and organisms containing them | |
| JPH11313678A (en) | Wasabi antibacterial protein gene | |
| WO2006066355A1 (en) | Chitin-binding peptides | |
| Ribeiro et al. | Plant antimicrobial peptides: From basic structures to applied research | |
| AU2012201612A1 (en) | Antifungal peptides | |
| AU2005215825A1 (en) | Antifungal peptides | |
| HU219505B (en) | Method for targeting plant derived intracellular proteins to the extracellular space and enhancing antipathogenic effect of such proteins |
Legal Events
| Date | Code | Title | Description |
|---|---|---|---|
| STCB | Information on status: application discontinuation |
Free format text: ABANDONED -- FAILURE TO RESPOND TO AN OFFICE ACTION |