EP1765992A2 - Verkürzte galnact2-polypeptide und nukleinsäuren - Google Patents
Verkürzte galnact2-polypeptide und nukleinsäurenInfo
- Publication number
- EP1765992A2 EP1765992A2 EP05758682A EP05758682A EP1765992A2 EP 1765992 A2 EP1765992 A2 EP 1765992A2 EP 05758682 A EP05758682 A EP 05758682A EP 05758682 A EP05758682 A EP 05758682A EP 1765992 A2 EP1765992 A2 EP 1765992A2
- Authority
- EP
- European Patent Office
- Prior art keywords
- polypeptide
- galnact2
- nucleic acid
- truncated
- leu
- Prior art date
- Legal status (The legal status is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the status listed.)
- Withdrawn
Links
- 108090000765 processed proteins & peptides Proteins 0.000 title claims abstract description 368
- 102000004196 processed proteins & peptides Human genes 0.000 title claims abstract description 328
- 229920001184 polypeptide Polymers 0.000 title claims abstract description 317
- 150000007523 nucleic acids Chemical class 0.000 title claims abstract description 183
- 108020004707 nucleic acids Proteins 0.000 title claims abstract description 138
- 102000039446 nucleic acids Human genes 0.000 title claims abstract description 138
- 238000000034 method Methods 0.000 claims abstract description 98
- 101000776623 Homo sapiens Chondroitin sulfate N-acetylgalactosaminyltransferase 2 Proteins 0.000 claims abstract description 71
- 210000004027 cell Anatomy 0.000 claims description 137
- 108090000623 proteins and genes Proteins 0.000 claims description 110
- 102000004169 proteins and genes Human genes 0.000 claims description 94
- OVRNDRQMDRJTHS-KEWYIRBNSA-N N-acetyl-D-galactosamine Chemical group CC(=O)N[C@H]1C(O)O[C@H](CO)[C@H](O)[C@@H]1O OVRNDRQMDRJTHS-KEWYIRBNSA-N 0.000 claims description 59
- 108091028043 Nucleic acid sequence Proteins 0.000 claims description 46
- 108010017080 Granulocyte Colony-Stimulating Factor Proteins 0.000 claims description 36
- 102000004269 Granulocyte Colony-Stimulating Factor Human genes 0.000 claims description 36
- 108010015899 Glycopeptides Proteins 0.000 claims description 31
- 102000002068 Glycopeptides Human genes 0.000 claims description 31
- 238000012546 transfer Methods 0.000 claims description 31
- 150000001720 carbohydrates Chemical class 0.000 claims description 24
- 125000003275 alpha amino acid group Chemical group 0.000 claims description 23
- 125000000539 amino acid group Chemical group 0.000 claims description 21
- DQJCDTNMLBYVAY-ZXXIYAEKSA-N (2S,5R,10R,13R)-16-{[(2R,3S,4R,5R)-3-{[(2S,3R,4R,5S,6R)-3-acetamido-4,5-dihydroxy-6-(hydroxymethyl)oxan-2-yl]oxy}-5-(ethylamino)-6-hydroxy-2-(hydroxymethyl)oxan-4-yl]oxy}-5-(4-aminobutyl)-10-carbamoyl-2,13-dimethyl-4,7,12,15-tetraoxo-3,6,11,14-tetraazaheptadecan-1-oic acid Chemical compound NCCCC[C@H](C(=O)N[C@@H](C)C(O)=O)NC(=O)CC[C@H](C(N)=O)NC(=O)[C@@H](C)NC(=O)C(C)O[C@@H]1[C@@H](NCC)C(O)O[C@H](CO)[C@H]1O[C@H]1[C@H](NC(C)=O)[C@@H](O)[C@H](O)[C@@H](CO)O1 DQJCDTNMLBYVAY-ZXXIYAEKSA-N 0.000 claims description 20
- 101710175625 Maltose/maltodextrin-binding periplasmic protein Proteins 0.000 claims description 16
- 229920001223 polyethylene glycol Polymers 0.000 claims description 16
- 241000588724 Escherichia coli Species 0.000 claims description 13
- 239000013604 expression vector Substances 0.000 claims description 13
- 239000002202 Polyethylene glycol Substances 0.000 claims description 12
- 102000003886 Glycoproteins Human genes 0.000 claims description 9
- 108090000288 Glycoproteins Proteins 0.000 claims description 9
- 102100022641 Coagulation factor IX Human genes 0.000 claims description 7
- 102000003951 Erythropoietin Human genes 0.000 claims description 7
- 108090000394 Erythropoietin Proteins 0.000 claims description 7
- 108010076282 Factor IX Proteins 0.000 claims description 7
- 229940105423 erythropoietin Drugs 0.000 claims description 7
- 229960004222 factor ix Drugs 0.000 claims description 7
- OXCMYAYHXIHQOA-UHFFFAOYSA-N potassium;[2-butyl-5-chloro-3-[[4-[2-(1,2,4-triaza-3-azanidacyclopenta-1,4-dien-5-yl)phenyl]phenyl]methyl]imidazol-4-yl]methanol Chemical compound [K+].CCCCC1=NC(Cl)=C(CO)N1CC1=CC=C(C=2C(=CC=CC=2)C2=N[N-]N=N2)C=C1 OXCMYAYHXIHQOA-UHFFFAOYSA-N 0.000 claims description 7
- 230000001105 regulatory effect Effects 0.000 claims description 7
- 230000004927 fusion Effects 0.000 claims description 6
- HNDVDQJCIGZPNO-UHFFFAOYSA-N histidine Natural products OC(=O)C(N)CC1=CN=CN1 HNDVDQJCIGZPNO-UHFFFAOYSA-N 0.000 claims description 6
- 210000004962 mammalian cell Anatomy 0.000 claims description 6
- 102000005720 Glutathione transferase Human genes 0.000 claims description 5
- 108010070675 Glutathione transferase Proteins 0.000 claims description 5
- 241000238631 Hexapoda Species 0.000 claims description 5
- 229920002472 Starch Polymers 0.000 claims description 5
- 235000019698 starch Nutrition 0.000 claims description 5
- 239000008107 starch Substances 0.000 claims description 5
- 241000255581 Drosophila <fruit fly, genus> Species 0.000 claims description 4
- 102000012673 Follicle Stimulating Hormone Human genes 0.000 claims description 4
- 108010079345 Follicle Stimulating Hormone Proteins 0.000 claims description 4
- 102000002265 Human Growth Hormone Human genes 0.000 claims description 4
- 108010000521 Human Growth Hormone Proteins 0.000 claims description 4
- 239000000854 Human Growth Hormone Substances 0.000 claims description 4
- 108010050904 Interferons Proteins 0.000 claims description 4
- 102000014150 Interferons Human genes 0.000 claims description 4
- 210000003527 eukaryotic cell Anatomy 0.000 claims description 4
- 229940028334 follicle stimulating hormone Drugs 0.000 claims description 4
- 230000002132 lysosomal effect Effects 0.000 claims description 4
- 108010002350 Interleukin-2 Proteins 0.000 claims description 3
- 102100020873 Interleukin-2 Human genes 0.000 claims description 3
- 230000002538 fungal effect Effects 0.000 claims description 3
- 229940047124 interferons Drugs 0.000 claims description 3
- 210000001236 prokaryotic cell Anatomy 0.000 claims description 3
- 125000003827 glycol group Chemical group 0.000 claims description 2
- 235000014469 Bacillus subtilis Nutrition 0.000 claims 1
- 239000013598 vector Substances 0.000 abstract description 61
- 239000000203 mixture Substances 0.000 abstract description 20
- 235000018102 proteins Nutrition 0.000 description 87
- FAPWRFPIFSIZLT-UHFFFAOYSA-M Sodium chloride Chemical compound [Na+].[Cl-] FAPWRFPIFSIZLT-UHFFFAOYSA-M 0.000 description 64
- 239000000370 acceptor Substances 0.000 description 60
- 230000000694 effects Effects 0.000 description 50
- 102000004190 Enzymes Human genes 0.000 description 49
- 108090000790 Enzymes Proteins 0.000 description 49
- 229940088598 enzyme Drugs 0.000 description 49
- 239000000499 gel Substances 0.000 description 42
- 125000003729 nucleotide group Chemical group 0.000 description 42
- MBLBDJOUHNCFQT-UHFFFAOYSA-N N-acetyl-D-galactosamine Natural products CC(=O)NC(C=O)C(O)C(O)C(O)CO MBLBDJOUHNCFQT-UHFFFAOYSA-N 0.000 description 40
- 238000003556 assay Methods 0.000 description 40
- 108700023372 Glycosyltransferases Proteins 0.000 description 37
- 108020004414 DNA Proteins 0.000 description 35
- 238000006243 chemical reaction Methods 0.000 description 34
- 239000011780 sodium chloride Substances 0.000 description 32
- 210000003000 inclusion body Anatomy 0.000 description 31
- 239000002773 nucleotide Substances 0.000 description 30
- 239000000872 buffer Substances 0.000 description 23
- 239000000047 product Substances 0.000 description 23
- 238000000746 purification Methods 0.000 description 23
- 102000051366 Glycosyltransferases Human genes 0.000 description 22
- 235000001014 amino acid Nutrition 0.000 description 22
- 229940024606 amino acid Drugs 0.000 description 21
- 150000002482 oligosaccharides Polymers 0.000 description 21
- KDXKERNSBIXSRK-UHFFFAOYSA-N Lysine Natural products NCCCCC(N)C(O)=O KDXKERNSBIXSRK-UHFFFAOYSA-N 0.000 description 20
- 108010066816 Polypeptide N-acetylgalactosaminyltransferase Proteins 0.000 description 20
- 150000001413 amino acids Chemical class 0.000 description 20
- 239000012564 Q sepharose fast flow resin Substances 0.000 description 19
- 229920001542 oligosaccharide Polymers 0.000 description 19
- 239000000523 sample Substances 0.000 description 18
- 102100039847 Globoside alpha-1,3-N-acetylgalactosaminyltransferase 1 Human genes 0.000 description 17
- 101000887519 Homo sapiens Globoside alpha-1,3-N-acetylgalactosaminyltransferase 1 Proteins 0.000 description 17
- LFTYTUAZOPRMMI-NESSUJCYSA-N UDP-N-acetyl-alpha-D-galactosamine Chemical compound O1[C@H](CO)[C@H](O)[C@H](O)[C@@H](NC(=O)C)[C@H]1O[P@](O)(=O)O[P@](O)(=O)OC[C@@H]1[C@@H](O)[C@@H](O)[C@H](N2C(NC(=O)C=C2)=O)O1 LFTYTUAZOPRMMI-NESSUJCYSA-N 0.000 description 17
- LFTYTUAZOPRMMI-UHFFFAOYSA-N UNPD164450 Natural products O1C(CO)C(O)C(O)C(NC(=O)C)C1OP(O)(=O)OP(O)(=O)OCC1C(O)C(O)C(N2C(NC(=O)C=C2)=O)O1 LFTYTUAZOPRMMI-UHFFFAOYSA-N 0.000 description 17
- 238000004519 manufacturing process Methods 0.000 description 16
- DHMQDGOQFOQNFH-UHFFFAOYSA-N Glycine Chemical compound NCC(O)=O DHMQDGOQFOQNFH-UHFFFAOYSA-N 0.000 description 15
- 108700014210 glycosyltransferase activity proteins Proteins 0.000 description 15
- 102000045442 glycosyltransferase activity proteins Human genes 0.000 description 15
- WHUUTDBJXJRKMK-UHFFFAOYSA-N Glutamic acid Natural products OC(=O)C(N)CCC(O)=O WHUUTDBJXJRKMK-UHFFFAOYSA-N 0.000 description 14
- 239000003814 drug Substances 0.000 description 14
- 238000010828 elution Methods 0.000 description 14
- QKNYBSVHEMOAJP-UHFFFAOYSA-N 2-amino-2-(hydroxymethyl)propane-1,3-diol;hydron;chloride Chemical compound Cl.OCC(N)(CO)CO QKNYBSVHEMOAJP-UHFFFAOYSA-N 0.000 description 13
- SQVRNKJHWKZAKO-UHFFFAOYSA-N beta-N-Acetyl-D-neuraminic acid Natural products CC(=O)NC1C(O)CC(O)(C(O)=O)OC1C(O)C(O)CO SQVRNKJHWKZAKO-UHFFFAOYSA-N 0.000 description 13
- 230000004071 biological effect Effects 0.000 description 13
- 230000004048 modification Effects 0.000 description 13
- 238000012986 modification Methods 0.000 description 13
- 239000008188 pellet Substances 0.000 description 13
- 239000013612 plasmid Substances 0.000 description 13
- 238000011160 research Methods 0.000 description 13
- XLYOFNOQVPJJNP-UHFFFAOYSA-N water Chemical compound O XLYOFNOQVPJJNP-UHFFFAOYSA-N 0.000 description 13
- 239000012634 fragment Substances 0.000 description 12
- 238000002835 absorbance Methods 0.000 description 11
- 102000037865 fusion proteins Human genes 0.000 description 11
- 108020001507 fusion proteins Proteins 0.000 description 11
- 230000006698 induction Effects 0.000 description 11
- 238000003752 polymerase chain reaction Methods 0.000 description 11
- 230000001225 therapeutic effect Effects 0.000 description 11
- 125000000729 N-terminal amino-acid group Chemical group 0.000 description 10
- 241000158500 Platanus racemosa Species 0.000 description 10
- 230000003197 catalytic effect Effects 0.000 description 10
- 238000000338 in vitro Methods 0.000 description 10
- 239000006166 lysate Substances 0.000 description 10
- 108020004999 messenger RNA Proteins 0.000 description 10
- 150000002772 monosaccharides Chemical class 0.000 description 10
- 239000011347 resin Substances 0.000 description 10
- 229920005989 resin Polymers 0.000 description 10
- 239000000243 solution Substances 0.000 description 10
- 241000701447 unidentified baculovirus Species 0.000 description 10
- 241000880493 Leptailurus serval Species 0.000 description 9
- 239000007983 Tris buffer Substances 0.000 description 9
- 108010076324 alanyl-glycyl-glycine Proteins 0.000 description 9
- 230000001580 bacterial effect Effects 0.000 description 9
- LENZDBCJOHFCAS-UHFFFAOYSA-N tris Chemical compound OCC(N)(CO)CO LENZDBCJOHFCAS-UHFFFAOYSA-N 0.000 description 9
- 230000003612 virological effect Effects 0.000 description 9
- DVWVZSJAYIJZFI-FXQIFTODSA-N Ala-Arg-Asn Chemical compound [H]N[C@@H](C)C(=O)N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CC(N)=O)C(O)=O DVWVZSJAYIJZFI-FXQIFTODSA-N 0.000 description 8
- XSQUKJJJFZCRTK-UHFFFAOYSA-N Urea Chemical compound NC(N)=O XSQUKJJJFZCRTK-UHFFFAOYSA-N 0.000 description 8
- 108010062796 arginyllysine Proteins 0.000 description 8
- 239000003153 chemical reaction reagent Substances 0.000 description 8
- 229920000642 polymer Polymers 0.000 description 8
- 239000011541 reaction mixture Substances 0.000 description 8
- 238000005063 solubilization Methods 0.000 description 8
- 230000007928 solubilization Effects 0.000 description 8
- 235000000346 sugar Nutrition 0.000 description 8
- 108091026890 Coding region Proteins 0.000 description 7
- KCXVZYZYPLLWCC-UHFFFAOYSA-N EDTA Chemical compound OC(=O)CN(CC(O)=O)CCN(CC(O)=O)CC(O)=O KCXVZYZYPLLWCC-UHFFFAOYSA-N 0.000 description 7
- DCXYFEDJOCDNAF-REOHCLBHSA-N L-asparagine Chemical compound OC(=O)[C@@H](N)CC(N)=O DCXYFEDJOCDNAF-REOHCLBHSA-N 0.000 description 7
- 239000007987 MES buffer Substances 0.000 description 7
- 108010066427 N-valyltryptophan Proteins 0.000 description 7
- 238000010367 cloning Methods 0.000 description 7
- BPHPUYQFMNQIOC-NXRLNHOXSA-N isopropyl beta-D-thiogalactopyranoside Chemical compound CC(C)S[C@@H]1O[C@H](CO)[C@H](O)[C@H](O)[C@H]1O BPHPUYQFMNQIOC-NXRLNHOXSA-N 0.000 description 7
- 239000003550 marker Substances 0.000 description 7
- 239000000463 material Substances 0.000 description 7
- 238000010369 molecular cloning Methods 0.000 description 7
- 239000013641 positive control Substances 0.000 description 7
- 238000002360 preparation method Methods 0.000 description 7
- 238000007420 radioactive assay Methods 0.000 description 7
- SQVRNKJHWKZAKO-OQPLDHBCSA-N sialic acid Chemical compound CC(=O)N[C@@H]1[C@@H](O)C[C@@](O)(C(O)=O)OC1[C@H](O)[C@H](O)CO SQVRNKJHWKZAKO-OQPLDHBCSA-N 0.000 description 7
- 238000002415 sodium dodecyl sulfate polyacrylamide gel electrophoresis Methods 0.000 description 7
- 239000000725 suspension Substances 0.000 description 7
- 108010070543 CMP-N-acetylneuraminate-alpha-N-acetylgalactosaminide alpha-2,6-sialyltransferase Proteins 0.000 description 6
- 239000004471 Glycine Substances 0.000 description 6
- LZDNBBYBDGBADK-UHFFFAOYSA-N L-valyl-L-tryptophan Natural products C1=CC=C2C(CC(NC(=O)C(N)C(C)C)C(O)=O)=CNC2=C1 LZDNBBYBDGBADK-UHFFFAOYSA-N 0.000 description 6
- TWRXJAOTZQYOKJ-UHFFFAOYSA-L Magnesium chloride Chemical compound [Mg+2].[Cl-].[Cl-] TWRXJAOTZQYOKJ-UHFFFAOYSA-L 0.000 description 6
- SITLTJHOQZFJGG-UHFFFAOYSA-N N-L-alpha-glutamyl-L-valine Natural products CC(C)C(C(O)=O)NC(=O)C(N)CCC(O)=O SITLTJHOQZFJGG-UHFFFAOYSA-N 0.000 description 6
- HEMHJVSKTPXQMS-UHFFFAOYSA-M Sodium hydroxide Chemical compound [OH-].[Na+] HEMHJVSKTPXQMS-UHFFFAOYSA-M 0.000 description 6
- 230000015572 biosynthetic process Effects 0.000 description 6
- 125000000837 carbohydrate group Chemical group 0.000 description 6
- -1 e.g. Proteins 0.000 description 6
- 238000002474 experimental method Methods 0.000 description 6
- 230000013595 glycosylation Effects 0.000 description 6
- 238000006206 glycosylation reaction Methods 0.000 description 6
- 108010026364 glycyl-glycyl-leucine Proteins 0.000 description 6
- 238000002955 isolation Methods 0.000 description 6
- MTCFGRXMJLQNBG-REOHCLBHSA-N (2S)-2-Amino-3-hydroxypropansäure Chemical compound OC[C@H](N)C(O)=O MTCFGRXMJLQNBG-REOHCLBHSA-N 0.000 description 5
- VGPWRRFOPXVGOH-BYPYZUCNSA-N Ala-Gly-Gly Chemical compound C[C@H](N)C(=O)NCC(=O)NCC(O)=O VGPWRRFOPXVGOH-BYPYZUCNSA-N 0.000 description 5
- UISQLSIBJKEJSS-GUBZILKMSA-N Arg-Arg-Ser Chemical compound NC(N)=NCCC[C@H](N)C(=O)N[C@@H](CCCN=C(N)N)C(=O)N[C@@H](CO)C(O)=O UISQLSIBJKEJSS-GUBZILKMSA-N 0.000 description 5
- JQFZHHSQMKZLRU-IUCAKERBSA-N Arg-Lys Chemical compound NCCCC[C@@H](C(O)=O)NC(=O)[C@@H](N)CCCN=C(N)N JQFZHHSQMKZLRU-IUCAKERBSA-N 0.000 description 5
- VYZBPPBKFCHCIS-WPRPVWTQSA-N Arg-Val-Gly Chemical compound OC(=O)CNC(=O)[C@H](C(C)C)NC(=O)[C@@H](N)CCCN=C(N)N VYZBPPBKFCHCIS-WPRPVWTQSA-N 0.000 description 5
- MDDXKBHIMYYJLW-FXQIFTODSA-N Asn-Met-Asp Chemical compound CSCC[C@@H](C(=O)N[C@@H](CC(=O)O)C(=O)O)NC(=O)[C@H](CC(=O)N)N MDDXKBHIMYYJLW-FXQIFTODSA-N 0.000 description 5
- 108091033380 Coding strand Proteins 0.000 description 5
- ZQNCUVODKOBSSO-XEGUGMAKSA-N Glu-Trp-Ala Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N[C@@H](CC1=CNC2=C1C=CC=C2)C(=O)N[C@@H](C)C(O)=O ZQNCUVODKOBSSO-XEGUGMAKSA-N 0.000 description 5
- PPBKJAQJAUHZKX-SRVKXCTJSA-N Leu-Cys-Leu Chemical compound CC(C)C[C@H](N)C(=O)N[C@@H](CS)C(=O)N[C@H](C(O)=O)CC(C)C PPBKJAQJAUHZKX-SRVKXCTJSA-N 0.000 description 5
- DAYQSYGBCUKVKT-VOAKCMCISA-N Leu-Thr-Lys Chemical compound CC(C)C[C@H](N)C(=O)N[C@@H]([C@@H](C)O)C(=O)N[C@@H](CCCCN)C(O)=O DAYQSYGBCUKVKT-VOAKCMCISA-N 0.000 description 5
- ATIPDCIQTUXABX-UWVGGRQHSA-N Lys-Leu Chemical compound CC(C)C[C@@H](C(O)=O)NC(=O)[C@@H](N)CCCCN ATIPDCIQTUXABX-UWVGGRQHSA-N 0.000 description 5
- 239000004472 Lysine Substances 0.000 description 5
- 229910021380 Manganese Chloride Inorganic materials 0.000 description 5
- GLFNIEUTAYBVOC-UHFFFAOYSA-L Manganese chloride Chemical compound Cl[Mn]Cl GLFNIEUTAYBVOC-UHFFFAOYSA-L 0.000 description 5
- HQVPQXMCQKXARZ-FXQIFTODSA-N Pro-Cys-Ser Chemical compound C1C[C@H](NC1)C(=O)N[C@@H](CS)C(=O)N[C@@H](CO)C(=O)O HQVPQXMCQKXARZ-FXQIFTODSA-N 0.000 description 5
- VYVBSMCZNHOZGD-RCWTZXSCSA-N Thr-Val-Val Chemical compound [H]N[C@@H]([C@@H](C)O)C(=O)N[C@@H](C(C)C)C(=O)N[C@@H](C(C)C)C(O)=O VYVBSMCZNHOZGD-RCWTZXSCSA-N 0.000 description 5
- 229920004890 Triton X-100 Polymers 0.000 description 5
- 239000013504 Triton X-100 Substances 0.000 description 5
- 108010064997 VPY tripeptide Proteins 0.000 description 5
- ZLFHAAGHGQBQQN-GUBZILKMSA-N Val-Ala-Pro Natural products CC(C)[C@H](N)C(=O)N[C@@H](C)C(=O)N1CCC[C@H]1C(O)=O ZLFHAAGHGQBQQN-GUBZILKMSA-N 0.000 description 5
- OQWNEUXPKHIEJO-NRPADANISA-N Val-Glu-Ser Chemical compound CC(C)[C@@H](C(=O)N[C@@H](CCC(=O)O)C(=O)N[C@@H](CO)C(=O)O)N OQWNEUXPKHIEJO-NRPADANISA-N 0.000 description 5
- 241000700605 Viruses Species 0.000 description 5
- 239000002299 complementary DNA Substances 0.000 description 5
- 238000000502 dialysis Methods 0.000 description 5
- 229940079593 drug Drugs 0.000 description 5
- 108010042598 glutamyl-aspartyl-glycine Proteins 0.000 description 5
- 108010020688 glycylhistidine Proteins 0.000 description 5
- 108010050848 glycylleucine Proteins 0.000 description 5
- 238000011534 incubation Methods 0.000 description 5
- 239000011565 manganese chloride Substances 0.000 description 5
- 108091033319 polynucleotide Proteins 0.000 description 5
- 102000040430 polynucleotide Human genes 0.000 description 5
- 239000002157 polynucleotide Substances 0.000 description 5
- 235000010482 polyoxyethylene sorbitan monooleate Nutrition 0.000 description 5
- 229920000053 polysorbate 80 Polymers 0.000 description 5
- 238000012545 processing Methods 0.000 description 5
- 108010070643 prolylglutamic acid Proteins 0.000 description 5
- 108010048397 seryl-lysyl-leucine Proteins 0.000 description 5
- 239000000126 substance Substances 0.000 description 5
- 239000006228 supernatant Substances 0.000 description 5
- 238000003786 synthesis reaction Methods 0.000 description 5
- 108010031491 threonyl-lysyl-glutamic acid Proteins 0.000 description 5
- 238000001890 transfection Methods 0.000 description 5
- 230000009466 transformation Effects 0.000 description 5
- 108010036387 trimethionine Proteins 0.000 description 5
- 108091032973 (ribonucleotides)n+m Proteins 0.000 description 4
- YLTKNGYYPIWKHZ-ACZMJKKPSA-N Ala-Ala-Glu Chemical compound C[C@H](N)C(=O)N[C@@H](C)C(=O)N[C@H](C(O)=O)CCC(O)=O YLTKNGYYPIWKHZ-ACZMJKKPSA-N 0.000 description 4
- DCVYRWFAMZFSDA-ZLUOBGJFSA-N Ala-Ser-Ala Chemical compound [H]N[C@@H](C)C(=O)N[C@@H](CO)C(=O)N[C@@H](C)C(O)=O DCVYRWFAMZFSDA-ZLUOBGJFSA-N 0.000 description 4
- BMNVSPMWMICFRV-DCAQKATOSA-N Arg-His-Asp Chemical compound NC(N)=NCCC[C@H](N)C(=O)N[C@H](C(=O)N[C@@H](CC(O)=O)C(O)=O)CC1=CN=CN1 BMNVSPMWMICFRV-DCAQKATOSA-N 0.000 description 4
- DGFXIWKPTDKBLF-AVGNSLFASA-N Arg-His-Val Chemical compound CC(C)[C@@H](C(=O)O)NC(=O)[C@H](CC1=CN=CN1)NC(=O)[C@H](CCCN=C(N)N)N DGFXIWKPTDKBLF-AVGNSLFASA-N 0.000 description 4
- AFNHFVVOJZBIJD-GUBZILKMSA-N Arg-Met-Asp Chemical compound [H]N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CCSC)C(=O)N[C@@H](CC(O)=O)C(O)=O AFNHFVVOJZBIJD-GUBZILKMSA-N 0.000 description 4
- IVPNEDNYYYFAGI-GARJFASQSA-N Asp-Leu-Pro Chemical compound CC(C)C[C@@H](C(=O)N1CCC[C@@H]1C(=O)O)NC(=O)[C@H](CC(=O)O)N IVPNEDNYYYFAGI-GARJFASQSA-N 0.000 description 4
- GKWFMNNNYZHJHV-SRVKXCTJSA-N Asp-Lys-Leu Chemical compound CC(C)C[C@@H](C(O)=O)NC(=O)[C@H](CCCCN)NC(=O)[C@@H](N)CC(O)=O GKWFMNNNYZHJHV-SRVKXCTJSA-N 0.000 description 4
- 125000001433 C-terminal amino-acid group Chemical group 0.000 description 4
- DZLQXIFVQFTFJY-BYPYZUCNSA-N Cys-Gly-Gly Chemical compound SC[C@H](N)C(=O)NCC(=O)NCC(O)=O DZLQXIFVQFTFJY-BYPYZUCNSA-N 0.000 description 4
- KBKGRMNVKPSQIF-XDTLVQLUSA-N Glu-Ala-Tyr Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N[C@@H](C)C(=O)N[C@@H](CC1=CC=C(O)C=C1)C(O)=O KBKGRMNVKPSQIF-XDTLVQLUSA-N 0.000 description 4
- UHVIQGKBMXEVGN-WDSKDSINSA-N Glu-Gly-Asn Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)NCC(=O)N[C@@H](CC(N)=O)C(O)=O UHVIQGKBMXEVGN-WDSKDSINSA-N 0.000 description 4
- VSRCAOIHMGCIJK-SRVKXCTJSA-N Glu-Leu-Arg Chemical compound OC(=O)CC[C@H](N)C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](CCCN=C(N)N)C(O)=O VSRCAOIHMGCIJK-SRVKXCTJSA-N 0.000 description 4
- RBXSZQRSEGYDFG-GUBZILKMSA-N Glu-Lys-Ser Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N[C@@H](CCCCN)C(=O)N[C@@H](CO)C(O)=O RBXSZQRSEGYDFG-GUBZILKMSA-N 0.000 description 4
- 108010053070 Glutathione Disulfide Proteins 0.000 description 4
- XRTDOIOIBMAXCT-NKWVEPMBSA-N Gly-Asn-Pro Chemical compound C1C[C@@H](N(C1)C(=O)[C@H](CC(=O)N)NC(=O)CN)C(=O)O XRTDOIOIBMAXCT-NKWVEPMBSA-N 0.000 description 4
- YZACQYVWLCQWBT-BQBZGAKWSA-N Gly-Cys-Arg Chemical compound [H]NCC(=O)N[C@@H](CS)C(=O)N[C@@H](CCCNC(N)=N)C(O)=O YZACQYVWLCQWBT-BQBZGAKWSA-N 0.000 description 4
- FKESCSGWBPUTPN-FOHZUACHSA-N Gly-Thr-Asn Chemical compound [H]NCC(=O)N[C@@H]([C@@H](C)O)C(=O)N[C@@H](CC(N)=O)C(O)=O FKESCSGWBPUTPN-FOHZUACHSA-N 0.000 description 4
- 102100039619 Granulocyte colony-stimulating factor Human genes 0.000 description 4
- CHIAUHSHDARFBD-ULQDDVLXSA-N His-Pro-Tyr Chemical compound C([C@H](N)C(=O)N1[C@@H](CCC1)C(=O)N[C@@H](CC=1C=CC(O)=CC=1)C(O)=O)C1=CN=CN1 CHIAUHSHDARFBD-ULQDDVLXSA-N 0.000 description 4
- GGXUJBKENKVYNV-ULQDDVLXSA-N His-Val-Phe Chemical compound CC(C)[C@@H](C(=O)N[C@@H](CC1=CC=CC=C1)C(=O)O)NC(=O)[C@H](CC2=CN=CN2)N GGXUJBKENKVYNV-ULQDDVLXSA-N 0.000 description 4
- 101000746367 Homo sapiens Granulocyte colony-stimulating factor Proteins 0.000 description 4
- AYFVYJQAPQTCCC-GBXIJSLDSA-N L-threonine Chemical compound C[C@@H](O)[C@H](N)C(O)=O AYFVYJQAPQTCCC-GBXIJSLDSA-N 0.000 description 4
- PVMPDMIKUVNOBD-CIUDSAMLSA-N Leu-Asp-Ser Chemical compound CC(C)C[C@H](N)C(=O)N[C@@H](CC(O)=O)C(=O)N[C@@H](CO)C(O)=O PVMPDMIKUVNOBD-CIUDSAMLSA-N 0.000 description 4
- FYPWFNKQVVEELI-ULQDDVLXSA-N Leu-Phe-Val Chemical compound CC(C)C[C@H](N)C(=O)N[C@H](C(=O)N[C@@H](C(C)C)C(O)=O)CC1=CC=CC=C1 FYPWFNKQVVEELI-ULQDDVLXSA-N 0.000 description 4
- CGHXMODRYJISSK-NHCYSSNCSA-N Leu-Val-Asp Chemical compound CC(C)C[C@H](N)C(=O)N[C@@H](C(C)C)C(=O)N[C@H](C(O)=O)CC(O)=O CGHXMODRYJISSK-NHCYSSNCSA-N 0.000 description 4
- UGTZHPSKYRIGRJ-YUMQZZPRSA-N Lys-Glu Chemical compound NCCCC[C@H](N)C(=O)N[C@H](C(O)=O)CCC(O)=O UGTZHPSKYRIGRJ-YUMQZZPRSA-N 0.000 description 4
- BDFHWFUAQLIMJO-KXNHARMFSA-N Lys-Thr-Pro Chemical compound C[C@H]([C@@H](C(=O)N1CCC[C@@H]1C(=O)O)NC(=O)[C@H](CCCCN)N)O BDFHWFUAQLIMJO-KXNHARMFSA-N 0.000 description 4
- XGZDDOKIHSYHTO-SZMVWBNQSA-N Lys-Trp-Glu Chemical compound C1=CC=C2C(C[C@H](NC(=O)[C@@H](N)CCCCN)C(=O)N[C@@H](CCC(O)=O)C(O)=O)=CNC2=C1 XGZDDOKIHSYHTO-SZMVWBNQSA-N 0.000 description 4
- SQVRNKJHWKZAKO-PFQGKNLYSA-N N-acetyl-beta-neuraminic acid Chemical compound CC(=O)N[C@@H]1[C@@H](O)C[C@@](O)(C(O)=O)O[C@H]1[C@H](O)[C@H](O)CO SQVRNKJHWKZAKO-PFQGKNLYSA-N 0.000 description 4
- WZEWCHQHNCMBEN-PMVMPFDFSA-N Phe-Lys-Trp Chemical compound C1=CC=C(C=C1)C[C@@H](C(=O)N[C@@H](CCCCN)C(=O)N[C@@H](CC2=CNC3=CC=CC=C32)C(=O)O)N WZEWCHQHNCMBEN-PMVMPFDFSA-N 0.000 description 4
- SFECXGVELZFBFJ-VEVYYDQMSA-N Pro-Asp-Thr Chemical compound [H]N1CCC[C@H]1C(=O)N[C@@H](CC(O)=O)C(=O)N[C@@H]([C@@H](C)O)C(O)=O SFECXGVELZFBFJ-VEVYYDQMSA-N 0.000 description 4
- 239000012614 Q-Sepharose Substances 0.000 description 4
- WDXYVIIVDIDOSX-DCAQKATOSA-N Ser-Arg-Leu Chemical compound CC(C)C[C@@H](C(O)=O)NC(=O)[C@@H](NC(=O)[C@@H](N)CO)CCCN=C(N)N WDXYVIIVDIDOSX-DCAQKATOSA-N 0.000 description 4
- ZIFYDQAFEMIZII-GUBZILKMSA-N Ser-Leu-Glu Chemical compound [H]N[C@@H](CO)C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](CCC(O)=O)C(O)=O ZIFYDQAFEMIZII-GUBZILKMSA-N 0.000 description 4
- XUDRHBPSPAPDJP-SRVKXCTJSA-N Ser-Lys-Leu Chemical compound CC(C)C[C@@H](C(O)=O)NC(=O)[C@H](CCCCN)NC(=O)[C@@H](N)CO XUDRHBPSPAPDJP-SRVKXCTJSA-N 0.000 description 4
- JAWGSPUJAXYXJA-IHRRRGAJSA-N Ser-Phe-Arg Chemical compound NC(N)=NCCC[C@@H](C(O)=O)NC(=O)[C@@H](NC(=O)[C@H](CO)N)CC1=CC=CC=C1 JAWGSPUJAXYXJA-IHRRRGAJSA-N 0.000 description 4
- XYEXCEPTALHNEV-RCWTZXSCSA-N Thr-Arg-Arg Chemical compound [H]N[C@@H]([C@@H](C)O)C(=O)N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CCCNC(N)=N)C(O)=O XYEXCEPTALHNEV-RCWTZXSCSA-N 0.000 description 4
- IMDMLDSVUSMAEJ-HJGDQZAQSA-N Thr-Leu-Asn Chemical compound [H]N[C@@H]([C@@H](C)O)C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](CC(N)=O)C(O)=O IMDMLDSVUSMAEJ-HJGDQZAQSA-N 0.000 description 4
- GYUUYCIXELGTJS-MEYUZBJRSA-N Thr-Phe-His Chemical compound C[C@H]([C@@H](C(=O)N[C@@H](CC1=CC=CC=C1)C(=O)N[C@@H](CC2=CN=CN2)C(=O)O)N)O GYUUYCIXELGTJS-MEYUZBJRSA-N 0.000 description 4
- ICNFHVUVCNWUAB-SZMVWBNQSA-N Trp-Arg-Val Chemical compound CC(C)[C@@H](C(=O)O)NC(=O)[C@H](CCCN=C(N)N)NC(=O)[C@H](CC1=CNC2=CC=CC=C21)N ICNFHVUVCNWUAB-SZMVWBNQSA-N 0.000 description 4
- GRSCONMARGNYHA-PMVMPFDFSA-N Trp-Lys-Phe Chemical compound [H]N[C@@H](CC1=CNC2=C1C=CC=C2)C(=O)N[C@@H](CCCCN)C(=O)N[C@@H](CC1=CC=CC=C1)C(O)=O GRSCONMARGNYHA-PMVMPFDFSA-N 0.000 description 4
- NWEGIYMHTZXVBP-JSGCOSHPSA-N Tyr-Val-Gly Chemical compound [H]N[C@@H](CC1=CC=C(O)C=C1)C(=O)N[C@@H](C(C)C)C(=O)NCC(O)=O NWEGIYMHTZXVBP-JSGCOSHPSA-N 0.000 description 4
- ZLFHAAGHGQBQQN-AEJSXWLSSA-N Val-Ala-Pro Chemical compound C[C@@H](C(=O)N1CCC[C@@H]1C(=O)O)NC(=O)[C@H](C(C)C)N ZLFHAAGHGQBQQN-AEJSXWLSSA-N 0.000 description 4
- MLADEWAIYAPAAU-IHRRRGAJSA-N Val-Lys-His Chemical compound CC(C)[C@@H](C(=O)N[C@@H](CCCCN)C(=O)N[C@@H](CC1=CN=CN1)C(=O)O)N MLADEWAIYAPAAU-IHRRRGAJSA-N 0.000 description 4
- 239000011543 agarose gel Substances 0.000 description 4
- 108010069020 alanyl-prolyl-glycine Proteins 0.000 description 4
- 108010041407 alanylaspartic acid Proteins 0.000 description 4
- KOSRFJWDECSPRO-UHFFFAOYSA-N alpha-L-glutamyl-L-glutamic acid Natural products OC(=O)CCC(N)C(=O)NC(CCC(O)=O)C(O)=O KOSRFJWDECSPRO-UHFFFAOYSA-N 0.000 description 4
- 108010013835 arginine glutamate Proteins 0.000 description 4
- 108010040443 aspartyl-aspartic acid Proteins 0.000 description 4
- 239000004202 carbamide Substances 0.000 description 4
- 235000014633 carbohydrates Nutrition 0.000 description 4
- 239000013592 cell lysate Substances 0.000 description 4
- 238000005119 centrifugation Methods 0.000 description 4
- 125000000151 cysteine group Chemical group N[C@@H](CS)C(=O)* 0.000 description 4
- FSXRLASFHBWESK-UHFFFAOYSA-N dipeptide phenylalanyl-tyrosine Natural products C=1C=C(O)C=CC=1CC(C(O)=O)NC(=O)C(N)CC1=CC=CC=C1 FSXRLASFHBWESK-UHFFFAOYSA-N 0.000 description 4
- 238000002523 gelfiltration Methods 0.000 description 4
- YPZRWBKMTBYPTK-BJDJZHNGSA-N glutathione disulfide Chemical compound OC(=O)[C@@H](N)CCC(=O)N[C@H](C(=O)NCC(O)=O)CSSC[C@@H](C(=O)NCC(O)=O)NC(=O)CC[C@H](N)C(O)=O YPZRWBKMTBYPTK-BJDJZHNGSA-N 0.000 description 4
- 125000003147 glycosyl group Chemical group 0.000 description 4
- VPZXBVLAVMBEQI-UHFFFAOYSA-N glycyl-DL-alpha-alanine Natural products OC(=O)C(C)NC(=O)CN VPZXBVLAVMBEQI-UHFFFAOYSA-N 0.000 description 4
- 108010000434 glycyl-alanyl-leucine Proteins 0.000 description 4
- 108010089804 glycyl-threonine Proteins 0.000 description 4
- 239000001963 growth medium Substances 0.000 description 4
- 230000036541 health Effects 0.000 description 4
- 229910052588 hydroxylapatite Inorganic materials 0.000 description 4
- 229930027917 kanamycin Natural products 0.000 description 4
- 229960000318 kanamycin Drugs 0.000 description 4
- SBUJHOSQTJFQJX-NOAMYHISSA-N kanamycin Chemical compound O[C@@H]1[C@@H](O)[C@H](O)[C@@H](CN)O[C@@H]1O[C@H]1[C@H](O)[C@@H](O[C@@H]2[C@@H]([C@@H](N)[C@H](O)[C@@H](CO)O2)O)[C@H](N)C[C@@H]1N SBUJHOSQTJFQJX-NOAMYHISSA-N 0.000 description 4
- 229930182823 kanamycin A Natural products 0.000 description 4
- 108010030617 leucyl-phenylalanyl-valine Proteins 0.000 description 4
- 108010009298 lysylglutamic acid Proteins 0.000 description 4
- 238000000816 matrix-assisted laser desorption--ionisation Methods 0.000 description 4
- 239000002609 medium Substances 0.000 description 4
- XYJRXVWERLGGKC-UHFFFAOYSA-D pentacalcium;hydroxide;triphosphate Chemical compound [OH-].[Ca+2].[Ca+2].[Ca+2].[Ca+2].[Ca+2].[O-]P([O-])([O-])=O.[O-]P([O-])([O-])=O.[O-]P([O-])([O-])=O XYJRXVWERLGGKC-UHFFFAOYSA-D 0.000 description 4
- 230000008569 process Effects 0.000 description 4
- 108010020755 prolyl-glycyl-glycine Proteins 0.000 description 4
- 108091008146 restriction endonucleases Proteins 0.000 description 4
- 239000007858 starting material Substances 0.000 description 4
- 238000013518 transcription Methods 0.000 description 4
- 230000035897 transcription Effects 0.000 description 4
- 108010015666 tryptophyl-leucyl-glutamic acid Proteins 0.000 description 4
- AXAVXPMQTGXXJZ-UHFFFAOYSA-N 2-aminoacetic acid;2-amino-2-(hydroxymethyl)propane-1,3-diol Chemical compound NCC(O)=O.OCC(N)(CO)CO AXAVXPMQTGXXJZ-UHFFFAOYSA-N 0.000 description 3
- MDNAVFBZPROEHO-DCAQKATOSA-N Ala-Lys-Val Chemical compound [H]N[C@@H](C)C(=O)N[C@@H](CCCCN)C(=O)N[C@@H](C(C)C)C(O)=O MDNAVFBZPROEHO-DCAQKATOSA-N 0.000 description 3
- MDNAVFBZPROEHO-UHFFFAOYSA-N Ala-Lys-Val Natural products CC(C)C(C(O)=O)NC(=O)C(NC(=O)C(C)N)CCCCN MDNAVFBZPROEHO-UHFFFAOYSA-N 0.000 description 3
- DHBKYZYFEXXUAK-ONGXEEELSA-N Ala-Phe-Gly Chemical compound OC(=O)CNC(=O)[C@@H](NC(=O)[C@@H](N)C)CC1=CC=CC=C1 DHBKYZYFEXXUAK-ONGXEEELSA-N 0.000 description 3
- KTXKIYXZQFWJKB-VZFHVOOUSA-N Ala-Thr-Ser Chemical compound [H]N[C@@H](C)C(=O)N[C@@H]([C@@H](C)O)C(=O)N[C@@H](CO)C(O)=O KTXKIYXZQFWJKB-VZFHVOOUSA-N 0.000 description 3
- 102100031969 Alpha-N-acetylgalactosaminide alpha-2,6-sialyltransferase 1 Human genes 0.000 description 3
- KXOPYFNQLVUOAQ-FXQIFTODSA-N Arg-Ser-Ala Chemical compound [H]N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CO)C(=O)N[C@@H](C)C(O)=O KXOPYFNQLVUOAQ-FXQIFTODSA-N 0.000 description 3
- SUMJNGAMIQSNGX-TUAOUCFPSA-N Arg-Val-Pro Chemical compound CC(C)[C@H](NC(=O)[C@@H](N)CCCNC(N)=N)C(=O)N1CCC[C@@H]1C(O)=O SUMJNGAMIQSNGX-TUAOUCFPSA-N 0.000 description 3
- 239000004475 Arginine Substances 0.000 description 3
- ULRPXVNMIIYDDJ-ACZMJKKPSA-N Asn-Glu-Ala Chemical compound C[C@@H](C(=O)O)NC(=O)[C@H](CCC(=O)O)NC(=O)[C@H](CC(=O)N)N ULRPXVNMIIYDDJ-ACZMJKKPSA-N 0.000 description 3
- RBOBTTLFPRSXKZ-BZSNNMDCSA-N Asn-Phe-Tyr Chemical compound [H]N[C@@H](CC(N)=O)C(=O)N[C@@H](CC1=CC=CC=C1)C(=O)N[C@@H](CC1=CC=C(O)C=C1)C(O)=O RBOBTTLFPRSXKZ-BZSNNMDCSA-N 0.000 description 3
- RRKCPMGSRIDLNC-AVGNSLFASA-N Asp-Glu-Tyr Chemical compound [H]N[C@@H](CC(O)=O)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CC1=CC=C(O)C=C1)C(O)=O RRKCPMGSRIDLNC-AVGNSLFASA-N 0.000 description 3
- CJUKAWUWBZCTDQ-SRVKXCTJSA-N Asp-Leu-Lys Chemical compound OC(=O)C[C@H](N)C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](CCCCN)C(O)=O CJUKAWUWBZCTDQ-SRVKXCTJSA-N 0.000 description 3
- GGRSYTUJHAZTFN-IHRRRGAJSA-N Asp-Pro-Tyr Chemical compound C1C[C@H](N(C1)C(=O)[C@H](CC(=O)O)N)C(=O)N[C@@H](CC2=CC=C(C=C2)O)C(=O)O GGRSYTUJHAZTFN-IHRRRGAJSA-N 0.000 description 3
- ZBYLEBZCVKLPCY-FXQIFTODSA-N Asp-Ser-Arg Chemical compound [H]N[C@@H](CC(O)=O)C(=O)N[C@@H](CO)C(=O)N[C@@H](CCCNC(N)=N)C(O)=O ZBYLEBZCVKLPCY-FXQIFTODSA-N 0.000 description 3
- NVXLFIPTHPKSKL-UBHSHLNASA-N Asp-Trp-Asn Chemical compound C1=CC=C2C(C[C@H](NC(=O)[C@H](CC(O)=O)N)C(=O)N[C@@H](CC(N)=O)C(O)=O)=CNC2=C1 NVXLFIPTHPKSKL-UBHSHLNASA-N 0.000 description 3
- CZIVKMOEXPILDK-SRVKXCTJSA-N Asp-Tyr-Ser Chemical compound [H]N[C@@H](CC(O)=O)C(=O)N[C@@H](CC1=CC=C(O)C=C1)C(=O)N[C@@H](CO)C(O)=O CZIVKMOEXPILDK-SRVKXCTJSA-N 0.000 description 3
- DCXYFEDJOCDNAF-UHFFFAOYSA-N Asparagine Natural products OC(=O)C(N)CC(N)=O DCXYFEDJOCDNAF-UHFFFAOYSA-N 0.000 description 3
- 102220547700 Carcinoembryonic antigen-related cell adhesion molecule 3_N74G_mutation Human genes 0.000 description 3
- 108020004705 Codon Proteins 0.000 description 3
- BLGNLNRBABWDST-CIUDSAMLSA-N Cys-Leu-Asp Chemical compound CC(C)C[C@@H](C(=O)N[C@@H](CC(=O)O)C(=O)O)NC(=O)[C@H](CS)N BLGNLNRBABWDST-CIUDSAMLSA-N 0.000 description 3
- 102000053602 DNA Human genes 0.000 description 3
- 108091029865 Exogenous DNA Proteins 0.000 description 3
- YKLNMGJYMNPBCP-ACZMJKKPSA-N Glu-Asn-Asp Chemical compound C(CC(=O)O)[C@@H](C(=O)N[C@@H](CC(=O)N)C(=O)N[C@@H](CC(=O)O)C(=O)O)N YKLNMGJYMNPBCP-ACZMJKKPSA-N 0.000 description 3
- WQZGKKKJIJFFOK-GASJEMHNSA-N Glucose Chemical compound OC[C@H]1OC(O)[C@H](O)[C@@H](O)[C@@H]1O WQZGKKKJIJFFOK-GASJEMHNSA-N 0.000 description 3
- HDNXXTBKOJKWNN-WDSKDSINSA-N Gly-Glu-Asn Chemical compound NCC(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CC(N)=O)C(O)=O HDNXXTBKOJKWNN-WDSKDSINSA-N 0.000 description 3
- KMSGYZQRXPUKGI-BYPYZUCNSA-N Gly-Gly-Asn Chemical compound NCC(=O)NCC(=O)N[C@H](C(O)=O)CC(N)=O KMSGYZQRXPUKGI-BYPYZUCNSA-N 0.000 description 3
- XPJBQTCXPJNIFE-ZETCQYMHSA-N Gly-Gly-Leu Chemical compound CC(C)C[C@@H](C(O)=O)NC(=O)CNC(=O)CN XPJBQTCXPJNIFE-ZETCQYMHSA-N 0.000 description 3
- YWAQATDNEKZFFK-BYPYZUCNSA-N Gly-Gly-Ser Chemical compound NCC(=O)NCC(=O)N[C@@H](CO)C(O)=O YWAQATDNEKZFFK-BYPYZUCNSA-N 0.000 description 3
- IALQAMYQJBZNSK-WHFBIAKZSA-N Gly-Ser-Asn Chemical compound [H]NCC(=O)N[C@@H](CO)C(=O)N[C@@H](CC(N)=O)C(O)=O IALQAMYQJBZNSK-WHFBIAKZSA-N 0.000 description 3
- PEDCQBHIVMGVHV-UHFFFAOYSA-N Glycerine Chemical compound OCC(O)CO PEDCQBHIVMGVHV-UHFFFAOYSA-N 0.000 description 3
- MAABHGXCIBEYQR-XVYDVKMFSA-N His-Asn-Ala Chemical compound C[C@@H](C(=O)O)NC(=O)[C@H](CC(=O)N)NC(=O)[C@H](CC1=CN=CN1)N MAABHGXCIBEYQR-XVYDVKMFSA-N 0.000 description 3
- 108090000604 Hydrolases Proteins 0.000 description 3
- 102000004157 Hydrolases Human genes 0.000 description 3
- 102100040018 Interferon alpha-2 Human genes 0.000 description 3
- 108010079944 Interferon-alpha2b Proteins 0.000 description 3
- UGTHTQWIQKEDEH-BQBZGAKWSA-N L-alanyl-L-prolylglycine zwitterion Chemical compound C[C@H](N)C(=O)N1CCC[C@H]1C(=O)NCC(O)=O UGTHTQWIQKEDEH-BQBZGAKWSA-N 0.000 description 3
- ODKSFYDXXFIFQN-BYPYZUCNSA-N L-arginine Chemical compound OC(=O)[C@@H](N)CCCN=C(N)N ODKSFYDXXFIFQN-BYPYZUCNSA-N 0.000 description 3
- CKLJMWTZIZZHCS-REOHCLBHSA-N L-aspartic acid Chemical compound OC(=O)[C@@H](N)CC(O)=O CKLJMWTZIZZHCS-REOHCLBHSA-N 0.000 description 3
- TYYLDKGBCJGJGW-UHFFFAOYSA-N L-tryptophan-L-tyrosine Natural products C=1NC2=CC=CC=C2C=1CC(N)C(=O)NC(C(O)=O)CC1=CC=C(O)C=C1 TYYLDKGBCJGJGW-UHFFFAOYSA-N 0.000 description 3
- OUYCCCASQSFEME-QMMMGPOBSA-N L-tyrosine Chemical compound OC(=O)[C@@H](N)CC1=CC=C(O)C=C1 OUYCCCASQSFEME-QMMMGPOBSA-N 0.000 description 3
- DSFYPIUSAMSERP-IHRRRGAJSA-N Leu-Leu-Arg Chemical compound CC(C)C[C@H](N)C(=O)N[C@@H](CC(C)C)C(=O)N[C@H](C(O)=O)CCCN=C(N)N DSFYPIUSAMSERP-IHRRRGAJSA-N 0.000 description 3
- YOKVEHGYYQEQOP-QWRGUYRKSA-N Leu-Leu-Gly Chemical compound CC(C)C[C@H](N)C(=O)N[C@@H](CC(C)C)C(=O)NCC(O)=O YOKVEHGYYQEQOP-QWRGUYRKSA-N 0.000 description 3
- ODRREERHVHMIPT-OEAJRASXSA-N Leu-Thr-Phe Chemical compound CC(C)C[C@H](N)C(=O)N[C@@H]([C@@H](C)O)C(=O)N[C@H](C(O)=O)CC1=CC=CC=C1 ODRREERHVHMIPT-OEAJRASXSA-N 0.000 description 3
- WBSCNDJQPKSPII-KKUMJFAQSA-N Lys-Lys-Lys Chemical compound NCCCC[C@H](N)C(=O)N[C@@H](CCCCN)C(=O)N[C@@H](CCCCN)C(O)=O WBSCNDJQPKSPII-KKUMJFAQSA-N 0.000 description 3
- ODTZHNZPINULEU-KKUMJFAQSA-N Lys-Phe-Asn Chemical compound C1=CC=C(C=C1)C[C@@H](C(=O)N[C@@H](CC(=O)N)C(=O)O)NC(=O)[C@H](CCCCN)N ODTZHNZPINULEU-KKUMJFAQSA-N 0.000 description 3
- OSOLWRWQADPDIQ-DCAQKATOSA-N Met-Asp-Leu Chemical compound [H]N[C@@H](CCSC)C(=O)N[C@@H](CC(O)=O)C(=O)N[C@@H](CC(C)C)C(O)=O OSOLWRWQADPDIQ-DCAQKATOSA-N 0.000 description 3
- WGBMNLCRYKSWAR-DCAQKATOSA-N Met-Asp-Lys Chemical compound CSCC[C@H](N)C(=O)N[C@@H](CC(O)=O)C(=O)N[C@H](C(O)=O)CCCCN WGBMNLCRYKSWAR-DCAQKATOSA-N 0.000 description 3
- NDJSSFWDYDUQID-YTWAJWBKSA-N Met-Thr-Pro Chemical compound C[C@H]([C@@H](C(=O)N1CCC[C@@H]1C(=O)O)NC(=O)[C@H](CCSC)N)O NDJSSFWDYDUQID-YTWAJWBKSA-N 0.000 description 3
- OVRNDRQMDRJTHS-UHFFFAOYSA-N N-acelyl-D-glucosamine Natural products CC(=O)NC1C(O)OC(CO)C(O)C1O OVRNDRQMDRJTHS-UHFFFAOYSA-N 0.000 description 3
- MBLBDJOUHNCFQT-LXGUWJNJSA-N N-acetylglucosamine Natural products CC(=O)N[C@@H](C=O)[C@@H](O)[C@H](O)[C@H](O)CO MBLBDJOUHNCFQT-LXGUWJNJSA-N 0.000 description 3
- SQVRNKJHWKZAKO-LUWBGTNYSA-N N-acetylneuraminic acid Chemical compound CC(=O)N[C@@H]1[C@@H](O)CC(O)(C(O)=O)O[C@H]1[C@H](O)[C@H](O)CO SQVRNKJHWKZAKO-LUWBGTNYSA-N 0.000 description 3
- KZNQNBZMBZJQJO-UHFFFAOYSA-N N-glycyl-L-proline Natural products NCC(=O)N1CCCC1C(O)=O KZNQNBZMBZJQJO-UHFFFAOYSA-N 0.000 description 3
- 108010079364 N-glycylalanine Proteins 0.000 description 3
- 108010087066 N2-tryptophyllysine Proteins 0.000 description 3
- FRMKIPSIZSFTTE-HJOGWXRNSA-N Phe-Tyr-Phe Chemical compound [H]N[C@@H](CC1=CC=CC=C1)C(=O)N[C@@H](CC1=CC=C(O)C=C1)C(=O)N[C@@H](CC1=CC=CC=C1)C(O)=O FRMKIPSIZSFTTE-HJOGWXRNSA-N 0.000 description 3
- IFMDQWDAJUMMJC-DCAQKATOSA-N Pro-Ala-Leu Chemical compound [H]N1CCC[C@H]1C(=O)N[C@@H](C)C(=O)N[C@@H](CC(C)C)C(O)=O IFMDQWDAJUMMJC-DCAQKATOSA-N 0.000 description 3
- 102000007056 Recombinant Fusion Proteins Human genes 0.000 description 3
- 108010008281 Recombinant Fusion Proteins Proteins 0.000 description 3
- 229920002684 Sepharose Polymers 0.000 description 3
- 108090000141 Sialyltransferases Proteins 0.000 description 3
- 102000003838 Sialyltransferases Human genes 0.000 description 3
- JMBRNXUOLJFURW-BEAPCOKYSA-N Thr-Phe-Pro Chemical compound C[C@H]([C@@H](C(=O)N[C@@H](CC1=CC=CC=C1)C(=O)N2CCC[C@@H]2C(=O)O)N)O JMBRNXUOLJFURW-BEAPCOKYSA-N 0.000 description 3
- AYFVYJQAPQTCCC-UHFFFAOYSA-N Threonine Natural products CC(O)C(N)C(O)=O AYFVYJQAPQTCCC-UHFFFAOYSA-N 0.000 description 3
- 239000004473 Threonine Substances 0.000 description 3
- 108020004566 Transfer RNA Proteins 0.000 description 3
- 102000004357 Transferases Human genes 0.000 description 3
- 108090000992 Transferases Proteins 0.000 description 3
- OFCKFBGRYHOKFP-IHPCNDPISA-N Trp-Asp-Tyr Chemical compound C1=CC=C2C(=C1)C(=CN2)C[C@@H](C(=O)N[C@@H](CC(=O)O)C(=O)N[C@@H](CC3=CC=C(C=C3)O)C(=O)O)N OFCKFBGRYHOKFP-IHPCNDPISA-N 0.000 description 3
- VCXWRWYFJLXITF-AUTRQRHGSA-N Tyr-Ala-Ala Chemical compound OC(=O)[C@H](C)NC(=O)[C@H](C)NC(=O)[C@@H](N)CC1=CC=C(O)C=C1 VCXWRWYFJLXITF-AUTRQRHGSA-N 0.000 description 3
- KSCVLGXNQXKUAR-JYJNAYRXSA-N Tyr-Leu-Glu Chemical compound [H]N[C@@H](CC1=CC=C(O)C=C1)C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](CCC(O)=O)C(O)=O KSCVLGXNQXKUAR-JYJNAYRXSA-N 0.000 description 3
- XQVRMLRMTAGSFJ-QXEWZRGKSA-N Val-Asp-Arg Chemical compound CC(C)[C@@H](C(=O)N[C@@H](CC(=O)O)C(=O)N[C@@H](CCCN=C(N)N)C(=O)O)N XQVRMLRMTAGSFJ-QXEWZRGKSA-N 0.000 description 3
- FPCIBLUVDNXPJO-XPUUQOCRSA-N Val-Cys-Gly Chemical compound CC(C)[C@H](N)C(=O)N[C@@H](CS)C(=O)NCC(O)=O FPCIBLUVDNXPJO-XPUUQOCRSA-N 0.000 description 3
- PIFJAFRUVWZRKR-QMMMGPOBSA-N Val-Gly-Gly Chemical compound CC(C)[C@H]([NH3+])C(=O)NCC(=O)NCC([O-])=O PIFJAFRUVWZRKR-QMMMGPOBSA-N 0.000 description 3
- MHHAWNPHDLCPLF-ULQDDVLXSA-N Val-Phe-Lys Chemical compound NCCCC[C@@H](C(O)=O)NC(=O)[C@@H](NC(=O)[C@@H](N)C(C)C)CC1=CC=CC=C1 MHHAWNPHDLCPLF-ULQDDVLXSA-N 0.000 description 3
- GBIUHAYJGWVNLN-UHFFFAOYSA-N Val-Ser-Pro Natural products CC(C)C(N)C(=O)NC(CO)C(=O)N1CCCC1C(O)=O GBIUHAYJGWVNLN-UHFFFAOYSA-N 0.000 description 3
- QHSSPPHOHJSTML-HOCLYGCPSA-N Val-Trp-Gly Chemical compound CC(C)[C@@H](C(=O)N[C@@H](CC1=CNC2=CC=CC=C21)C(=O)NCC(=O)O)N QHSSPPHOHJSTML-HOCLYGCPSA-N 0.000 description 3
- RFZFBOQPPFCOKG-BZSNNMDCSA-N Val-Trp-Met Chemical compound CC(C)[C@@H](C(=O)N[C@@H](CC1=CNC2=CC=CC=C21)C(=O)N[C@@H](CCSC)C(=O)O)N RFZFBOQPPFCOKG-BZSNNMDCSA-N 0.000 description 3
- LLJLBRRXKZTTRD-GUBZILKMSA-N Val-Val-Ser Chemical compound CC(C)[C@@H](C(=O)N[C@@H](C(C)C)C(=O)N[C@@H](CO)C(=O)O)N LLJLBRRXKZTTRD-GUBZILKMSA-N 0.000 description 3
- ODKSFYDXXFIFQN-UHFFFAOYSA-N arginine Natural products OC(=O)C(N)CCCNC(N)=N ODKSFYDXXFIFQN-UHFFFAOYSA-N 0.000 description 3
- 235000009697 arginine Nutrition 0.000 description 3
- 108010018691 arginyl-threonyl-arginine Proteins 0.000 description 3
- 235000009582 asparagine Nutrition 0.000 description 3
- 229960001230 asparagine Drugs 0.000 description 3
- 108010093581 aspartyl-proline Proteins 0.000 description 3
- MSWZFWKMSRAUBD-UHFFFAOYSA-N beta-D-galactosamine Natural products NC1C(O)OC(CO)C(O)C1O MSWZFWKMSRAUBD-UHFFFAOYSA-N 0.000 description 3
- 239000007795 chemical reaction product Substances 0.000 description 3
- 239000003795 chemical substances by application Substances 0.000 description 3
- 150000001875 compounds Chemical class 0.000 description 3
- 238000012217 deletion Methods 0.000 description 3
- 230000037430 deletion Effects 0.000 description 3
- 238000011161 development Methods 0.000 description 3
- 230000029087 digestion Effects 0.000 description 3
- 238000010790 dilution Methods 0.000 description 3
- 239000012895 dilution Substances 0.000 description 3
- 238000004520 electroporation Methods 0.000 description 3
- 230000002255 enzymatic effect Effects 0.000 description 3
- 230000002068 genetic effect Effects 0.000 description 3
- RWSXRVCMGQZWBV-WDSKDSINSA-N glutathione Chemical compound OC(=O)[C@@H](N)CCC(=O)N[C@@H](CS)C(=O)NCC(O)=O RWSXRVCMGQZWBV-WDSKDSINSA-N 0.000 description 3
- 150000004676 glycans Chemical class 0.000 description 3
- HPAIKDPJURGQLN-UHFFFAOYSA-N glycyl-L-histidyl-L-phenylalanine Natural products C=1C=CC=CC=1CC(C(O)=O)NC(=O)C(NC(=O)CN)CC1=CN=CN1 HPAIKDPJURGQLN-UHFFFAOYSA-N 0.000 description 3
- XKUKSGPZAADMRA-UHFFFAOYSA-N glycyl-glycyl-glycine Chemical compound NCC(=O)NCC(=O)NCC(O)=O XKUKSGPZAADMRA-UHFFFAOYSA-N 0.000 description 3
- 108010025306 histidylleucine Proteins 0.000 description 3
- 238000003780 insertion Methods 0.000 description 3
- 230000037431 insertion Effects 0.000 description 3
- OOYGSFOGFJDDHP-KMCOLRRFSA-N kanamycin A sulfate Chemical compound OS(O)(=O)=O.O[C@@H]1[C@@H](O)[C@H](O)[C@@H](CN)O[C@@H]1O[C@H]1[C@H](O)[C@@H](O[C@@H]2[C@@H]([C@@H](N)[C@H](O)[C@@H](CO)O2)O)[C@H](N)C[C@@H]1N OOYGSFOGFJDDHP-KMCOLRRFSA-N 0.000 description 3
- 229960002064 kanamycin sulfate Drugs 0.000 description 3
- 108010003700 lysyl aspartic acid Proteins 0.000 description 3
- 229910001629 magnesium chloride Inorganic materials 0.000 description 3
- 230000001404 mediated effect Effects 0.000 description 3
- 229950006780 n-acetylglucosamine Drugs 0.000 description 3
- 239000013642 negative control Substances 0.000 description 3
- 108010018625 phenylalanylarginine Proteins 0.000 description 3
- 229920001282 polysaccharide Polymers 0.000 description 3
- 239000005017 polysaccharide Substances 0.000 description 3
- 238000001742 protein purification Methods 0.000 description 3
- 230000002285 radioactive effect Effects 0.000 description 3
- 108010048818 seryl-histidine Proteins 0.000 description 3
- 125000005629 sialic acid group Chemical class 0.000 description 3
- 238000003756 stirring Methods 0.000 description 3
- 238000006467 substitution reaction Methods 0.000 description 3
- 239000000758 substrate Substances 0.000 description 3
- 230000001131 transforming effect Effects 0.000 description 3
- OUYCCCASQSFEME-UHFFFAOYSA-N tyrosine Natural products OC(=O)C(N)CC1=CC=C(O)C=C1 OUYCCCASQSFEME-UHFFFAOYSA-N 0.000 description 3
- 108010003137 tyrosyltyrosine Proteins 0.000 description 3
- WOJJIRYPFAZEPF-YFKPBYRVSA-N 2-[[(2s)-2-[[2-[(2-azaniumylacetyl)amino]acetyl]amino]propanoyl]amino]acetate Chemical compound OC(=O)CNC(=O)[C@H](C)NC(=O)CNC(=O)CN WOJJIRYPFAZEPF-YFKPBYRVSA-N 0.000 description 2
- TTXMOJWKNRJWQJ-FXQIFTODSA-N Ala-Arg-Ser Chemical compound OC[C@@H](C(O)=O)NC(=O)[C@@H](NC(=O)[C@@H](N)C)CCCN=C(N)N TTXMOJWKNRJWQJ-FXQIFTODSA-N 0.000 description 2
- KXEVYGKATAMXJJ-ACZMJKKPSA-N Ala-Glu-Asp Chemical compound C[C@H](N)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CC(O)=O)C(O)=O KXEVYGKATAMXJJ-ACZMJKKPSA-N 0.000 description 2
- NLYYHIKRBRMAJV-AEJSXWLSSA-N Ala-Val-Pro Chemical compound C[C@@H](C(=O)N[C@@H](C(C)C)C(=O)N1CCC[C@@H]1C(=O)O)N NLYYHIKRBRMAJV-AEJSXWLSSA-N 0.000 description 2
- ITVINTQUZMQWJR-QXEWZRGKSA-N Arg-Asn-Val Chemical compound [H]N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CC(N)=O)C(=O)N[C@@H](C(C)C)C(O)=O ITVINTQUZMQWJR-QXEWZRGKSA-N 0.000 description 2
- AQPVUEJJARLJHB-BQBZGAKWSA-N Arg-Gly-Ala Chemical compound OC(=O)[C@H](C)NC(=O)CNC(=O)[C@@H](N)CCCN=C(N)N AQPVUEJJARLJHB-BQBZGAKWSA-N 0.000 description 2
- CVXXSWQORBZAAA-SRVKXCTJSA-N Arg-Lys-Glu Chemical compound OC(=O)CC[C@@H](C(O)=O)NC(=O)[C@H](CCCCN)NC(=O)[C@@H](N)CCCN=C(N)N CVXXSWQORBZAAA-SRVKXCTJSA-N 0.000 description 2
- DNLQVHBBMPZUGJ-BQBZGAKWSA-N Arg-Ser-Gly Chemical compound [H]N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CO)C(=O)NCC(O)=O DNLQVHBBMPZUGJ-BQBZGAKWSA-N 0.000 description 2
- ASQKVGRCKOFKIU-KZVJFYERSA-N Arg-Thr-Ala Chemical compound C[C@H]([C@@H](C(=O)N[C@@H](C)C(=O)O)NC(=O)[C@H](CCCN=C(N)N)N)O ASQKVGRCKOFKIU-KZVJFYERSA-N 0.000 description 2
- AIFHRTPABBBHKU-RCWTZXSCSA-N Arg-Thr-Arg Chemical compound NC(N)=NCCC[C@H](N)C(=O)N[C@@H]([C@H](O)C)C(=O)N[C@@H](CCCN=C(N)N)C(O)=O AIFHRTPABBBHKU-RCWTZXSCSA-N 0.000 description 2
- PYDIIVKGTBRIEL-SZMVWBNQSA-N Arg-Trp-Pro Chemical compound [H]N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CC1=CNC2=C1C=CC=C2)C(=O)N1CCC[C@H]1C(O)=O PYDIIVKGTBRIEL-SZMVWBNQSA-N 0.000 description 2
- ZDOQDYFZNGASEY-BIIVOSGPSA-N Asn-Asp-Pro Chemical compound C1C[C@@H](N(C1)C(=O)[C@H](CC(=O)O)NC(=O)[C@H](CC(=O)N)N)C(=O)O ZDOQDYFZNGASEY-BIIVOSGPSA-N 0.000 description 2
- VYLVOMUVLMGCRF-ZLUOBGJFSA-N Asn-Asp-Ser Chemical compound NC(=O)C[C@H](N)C(=O)N[C@@H](CC(O)=O)C(=O)N[C@@H](CO)C(O)=O VYLVOMUVLMGCRF-ZLUOBGJFSA-N 0.000 description 2
- HDHZCEDPLTVHFZ-GUBZILKMSA-N Asn-Leu-Glu Chemical compound [H]N[C@@H](CC(N)=O)C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](CCC(O)=O)C(O)=O HDHZCEDPLTVHFZ-GUBZILKMSA-N 0.000 description 2
- NCFJQJRLQJEECD-NHCYSSNCSA-N Asn-Leu-Val Chemical compound [H]N[C@@H](CC(N)=O)C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](C(C)C)C(O)=O NCFJQJRLQJEECD-NHCYSSNCSA-N 0.000 description 2
- GHWWTICYPDKPTE-NGZCFLSTSA-N Asn-Val-Pro Chemical compound CC(C)[C@@H](C(=O)N1CCC[C@@H]1C(=O)O)NC(=O)[C@H](CC(=O)N)N GHWWTICYPDKPTE-NGZCFLSTSA-N 0.000 description 2
- HBUJSDCLZCXXCW-YDHLFZDLSA-N Asn-Val-Tyr Chemical compound NC(=O)C[C@H](N)C(=O)N[C@@H](C(C)C)C(=O)N[C@H](C(O)=O)CC1=CC=C(O)C=C1 HBUJSDCLZCXXCW-YDHLFZDLSA-N 0.000 description 2
- KRXIWXCXOARFNT-ZLUOBGJFSA-N Asp-Ala-Ala Chemical compound OC(=O)[C@H](C)NC(=O)[C@H](C)NC(=O)[C@@H](N)CC(O)=O KRXIWXCXOARFNT-ZLUOBGJFSA-N 0.000 description 2
- OERMIMJQPQUIPK-FXQIFTODSA-N Asp-Arg-Ala Chemical compound [H]N[C@@H](CC(O)=O)C(=O)N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](C)C(O)=O OERMIMJQPQUIPK-FXQIFTODSA-N 0.000 description 2
- DTNUIAJCPRMNBT-WHFBIAKZSA-N Asp-Gly-Ala Chemical compound [H]N[C@@H](CC(O)=O)C(=O)NCC(=O)N[C@@H](C)C(O)=O DTNUIAJCPRMNBT-WHFBIAKZSA-N 0.000 description 2
- WSGVTKZFVJSJOG-RCOVLWMOSA-N Asp-Gly-Val Chemical compound [H]N[C@@H](CC(O)=O)C(=O)NCC(=O)N[C@@H](C(C)C)C(O)=O WSGVTKZFVJSJOG-RCOVLWMOSA-N 0.000 description 2
- OEDJQRXNDRUGEU-SRVKXCTJSA-N Asp-Leu-His Chemical compound N[C@@H](CC(O)=O)C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](CC1=CNC=N1)C(=O)O OEDJQRXNDRUGEU-SRVKXCTJSA-N 0.000 description 2
- YWLDTBBUHZJQHW-KKUMJFAQSA-N Asp-Lys-Phe Chemical compound C1=CC=C(C=C1)C[C@@H](C(=O)O)NC(=O)[C@H](CCCCN)NC(=O)[C@H](CC(=O)O)N YWLDTBBUHZJQHW-KKUMJFAQSA-N 0.000 description 2
- WDMNFNXKGSLIOB-GUBZILKMSA-N Asp-Met-Met Chemical compound CSCC[C@@H](C(=O)N[C@@H](CCSC)C(=O)O)NC(=O)[C@H](CC(=O)O)N WDMNFNXKGSLIOB-GUBZILKMSA-N 0.000 description 2
- IDDMGSKZQDEDGA-SRVKXCTJSA-N Asp-Phe-Asn Chemical compound OC(=O)C[C@H](N)C(=O)N[C@H](C(=O)N[C@@H](CC(N)=O)C(O)=O)CC1=CC=CC=C1 IDDMGSKZQDEDGA-SRVKXCTJSA-N 0.000 description 2
- UKGGPJNBONZZCM-WDSKDSINSA-N Asp-Pro Chemical compound OC(=O)C[C@H](N)C(=O)N1CCC[C@H]1C(O)=O UKGGPJNBONZZCM-WDSKDSINSA-N 0.000 description 2
- AHWRSSLYSGLBGD-CIUDSAMLSA-N Asp-Pro-Glu Chemical compound OC(=O)C[C@H](N)C(=O)N1CCC[C@H]1C(=O)N[C@@H](CCC(O)=O)C(O)=O AHWRSSLYSGLBGD-CIUDSAMLSA-N 0.000 description 2
- 238000009631 Broth culture Methods 0.000 description 2
- KLLFLHBKSJAUMZ-ACZMJKKPSA-N Cys-Asn-Glu Chemical compound C(CC(=O)O)[C@@H](C(=O)O)NC(=O)[C@H](CC(=O)N)NC(=O)[C@H](CS)N KLLFLHBKSJAUMZ-ACZMJKKPSA-N 0.000 description 2
- XZKJEOMFLDVXJG-KATARQTJSA-N Cys-Leu-Thr Chemical compound C[C@H]([C@@H](C(=O)O)NC(=O)[C@H](CC(C)C)NC(=O)[C@H](CS)N)O XZKJEOMFLDVXJG-KATARQTJSA-N 0.000 description 2
- XMVZMBGFIOQONW-GARJFASQSA-N Cys-Lys-Pro Chemical compound C1C[C@@H](N(C1)C(=O)[C@H](CCCCN)NC(=O)[C@H](CS)N)C(=O)O XMVZMBGFIOQONW-GARJFASQSA-N 0.000 description 2
- AEMOLEFTQBMNLQ-AQKNRBDQSA-N D-glucopyranuronic acid Chemical compound OC1O[C@H](C(O)=O)[C@@H](O)[C@H](O)[C@H]1O AEMOLEFTQBMNLQ-AQKNRBDQSA-N 0.000 description 2
- WHUUTDBJXJRKMK-GSVOUGTGSA-N D-glutamic acid Chemical compound OC(=O)[C@H](N)CCC(O)=O WHUUTDBJXJRKMK-GSVOUGTGSA-N 0.000 description 2
- SHZGCJCMOBCMKK-UHFFFAOYSA-N D-mannomethylose Natural products CC1OC(O)C(O)C(O)C1O SHZGCJCMOBCMKK-UHFFFAOYSA-N 0.000 description 2
- SRBFZHDQGSBBOR-IOVATXLUSA-N D-xylopyranose Chemical compound O[C@@H]1COC(O)[C@H](O)[C@H]1O SRBFZHDQGSBBOR-IOVATXLUSA-N 0.000 description 2
- PNNNRSAQSRJVSB-SLPGGIOYSA-N Fucose Natural products C[C@H](O)[C@@H](O)[C@H](O)[C@H](O)C=O PNNNRSAQSRJVSB-SLPGGIOYSA-N 0.000 description 2
- IAJILQKETJEXLJ-UHFFFAOYSA-N Galacturonsaeure Natural products O=CC(O)C(O)C(O)C(O)C(O)=O IAJILQKETJEXLJ-UHFFFAOYSA-N 0.000 description 2
- 241000287828 Gallus gallus Species 0.000 description 2
- DYFJZDDQPNIPAB-NHCYSSNCSA-N Glu-Arg-Val Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](C(C)C)C(O)=O DYFJZDDQPNIPAB-NHCYSSNCSA-N 0.000 description 2
- PCBBLFVHTYNQGG-LAEOZQHASA-N Glu-Asn-Val Chemical compound CC(C)[C@@H](C(=O)O)NC(=O)[C@H](CC(=O)N)NC(=O)[C@H](CCC(=O)O)N PCBBLFVHTYNQGG-LAEOZQHASA-N 0.000 description 2
- DSPQRJXOIXHOHK-WDSKDSINSA-N Glu-Asp-Gly Chemical compound OC(=O)CC[C@H](N)C(=O)N[C@@H](CC(O)=O)C(=O)NCC(O)=O DSPQRJXOIXHOHK-WDSKDSINSA-N 0.000 description 2
- KLJMRPIBBLTDGE-ACZMJKKPSA-N Glu-Cys-Asn Chemical compound OC(=O)CC[C@H](N)C(=O)N[C@@H](CS)C(=O)N[C@@H](CC(N)=O)C(O)=O KLJMRPIBBLTDGE-ACZMJKKPSA-N 0.000 description 2
- MUSGDMDGNGXULI-DCAQKATOSA-N Glu-Glu-Leu Chemical compound CC(C)C[C@@H](C(O)=O)NC(=O)[C@H](CCC(O)=O)NC(=O)[C@@H](N)CCC(O)=O MUSGDMDGNGXULI-DCAQKATOSA-N 0.000 description 2
- ZGKXAUIVGIBISK-SZMVWBNQSA-N Glu-His-Trp Chemical compound N[C@@H](CCC(O)=O)C(=O)N[C@@H](Cc1c[nH]cn1)C(=O)N[C@@H](Cc1c[nH]c2ccccc12)C(O)=O ZGKXAUIVGIBISK-SZMVWBNQSA-N 0.000 description 2
- ATVYZJGOZLVXDK-IUCAKERBSA-N Glu-Leu-Gly Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N[C@@H](CC(C)C)C(=O)NCC(O)=O ATVYZJGOZLVXDK-IUCAKERBSA-N 0.000 description 2
- SUIAHERNFYRBDZ-GVXVVHGQSA-N Glu-Lys-Val Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N[C@@H](CCCCN)C(=O)N[C@@H](C(C)C)C(O)=O SUIAHERNFYRBDZ-GVXVVHGQSA-N 0.000 description 2
- MFYLRRCYBBJYPI-JYJNAYRXSA-N Glu-Tyr-Lys Chemical compound C1=CC(=CC=C1C[C@@H](C(=O)N[C@@H](CCCCN)C(=O)O)NC(=O)[C@H](CCC(=O)O)N)O MFYLRRCYBBJYPI-JYJNAYRXSA-N 0.000 description 2
- FUESBOMYALLFNI-VKHMYHEASA-N Gly-Asn Chemical compound NCC(=O)N[C@H](C(O)=O)CC(N)=O FUESBOMYALLFNI-VKHMYHEASA-N 0.000 description 2
- SOEATRRYCIPEHA-BQBZGAKWSA-N Gly-Glu-Glu Chemical compound [H]NCC(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CCC(O)=O)C(O)=O SOEATRRYCIPEHA-BQBZGAKWSA-N 0.000 description 2
- KAJAOGBVWCYGHZ-JTQLQIEISA-N Gly-Gly-Phe Chemical compound [NH3+]CC(=O)NCC(=O)N[C@H](C([O-])=O)CC1=CC=CC=C1 KAJAOGBVWCYGHZ-JTQLQIEISA-N 0.000 description 2
- NNCSJUBVFBDDLC-YUMQZZPRSA-N Gly-Leu-Ser Chemical compound NCC(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](CO)C(O)=O NNCSJUBVFBDDLC-YUMQZZPRSA-N 0.000 description 2
- CVFOYJJOZYYEPE-KBPBESRZSA-N Gly-Lys-Tyr Chemical compound [H]NCC(=O)N[C@@H](CCCCN)C(=O)N[C@@H](CC1=CC=C(O)C=C1)C(O)=O CVFOYJJOZYYEPE-KBPBESRZSA-N 0.000 description 2
- OQQKUTVULYLCDG-ONGXEEELSA-N Gly-Lys-Val Chemical compound CC(C)[C@H](NC(=O)[C@H](CCCCN)NC(=O)CN)C(O)=O OQQKUTVULYLCDG-ONGXEEELSA-N 0.000 description 2
- SOEGEPHNZOISMT-BYPYZUCNSA-N Gly-Ser-Gly Chemical compound NCC(=O)N[C@@H](CO)C(=O)NCC(O)=O SOEGEPHNZOISMT-BYPYZUCNSA-N 0.000 description 2
- CUVBTVWFVIIDOC-YEPSODPASA-N Gly-Thr-Val Chemical compound CC(C)[C@@H](C(O)=O)NC(=O)[C@H]([C@@H](C)O)NC(=O)CN CUVBTVWFVIIDOC-YEPSODPASA-N 0.000 description 2
- IZVICCORZOSGPT-JSGCOSHPSA-N Gly-Val-Tyr Chemical compound [H]NCC(=O)N[C@@H](C(C)C)C(=O)N[C@@H](CC1=CC=C(O)C=C1)C(O)=O IZVICCORZOSGPT-JSGCOSHPSA-N 0.000 description 2
- KSOBNUBCYHGUKH-UWVGGRQHSA-N Gly-Val-Val Chemical compound CC(C)[C@@H](C(O)=O)NC(=O)[C@H](C(C)C)NC(=O)CN KSOBNUBCYHGUKH-UWVGGRQHSA-N 0.000 description 2
- ZRALSGWEFCBTJO-UHFFFAOYSA-N Guanidine Chemical compound NC(N)=N ZRALSGWEFCBTJO-UHFFFAOYSA-N 0.000 description 2
- NYHBQMYGNKIUIF-UUOKFMHZSA-N Guanosine Chemical compound C1=NC=2C(=O)NC(N)=NC=2N1[C@@H]1O[C@H](CO)[C@@H](O)[C@H]1O NYHBQMYGNKIUIF-UUOKFMHZSA-N 0.000 description 2
- JBCLFWXMTIKCCB-UHFFFAOYSA-N H-Gly-Phe-OH Natural products NCC(=O)NC(C(O)=O)CC1=CC=CC=C1 JBCLFWXMTIKCCB-UHFFFAOYSA-N 0.000 description 2
- CYHWWHKRCKHYGQ-GUBZILKMSA-N His-Cys-Glu Chemical compound C1=C(NC=N1)C[C@@H](C(=O)N[C@@H](CS)C(=O)N[C@@H](CCC(=O)O)C(=O)O)N CYHWWHKRCKHYGQ-GUBZILKMSA-N 0.000 description 2
- CWSZWFILCNSNEX-CIUDSAMLSA-N His-Ser-Asn Chemical compound C1=C(NC=N1)C[C@@H](C(=O)N[C@@H](CO)C(=O)N[C@@H](CC(=O)N)C(=O)O)N CWSZWFILCNSNEX-CIUDSAMLSA-N 0.000 description 2
- CLRLHXKNIYJWAW-UHFFFAOYSA-N KDN Natural products OCC(O)C(O)C1OC(O)(C(O)=O)CC(O)C1O CLRLHXKNIYJWAW-UHFFFAOYSA-N 0.000 description 2
- QNAYBMKLOCPYGJ-REOHCLBHSA-N L-alanine Chemical compound C[C@H](N)C(O)=O QNAYBMKLOCPYGJ-REOHCLBHSA-N 0.000 description 2
- SHZGCJCMOBCMKK-DHVFOXMCSA-N L-fucopyranose Chemical compound C[C@@H]1OC(O)[C@@H](O)[C@H](O)[C@@H]1O SHZGCJCMOBCMKK-DHVFOXMCSA-N 0.000 description 2
- AEMOLEFTQBMNLQ-HNFCZKTMSA-N L-idopyranuronic acid Chemical compound OC1O[C@@H](C(O)=O)[C@@H](O)[C@H](O)[C@H]1O AEMOLEFTQBMNLQ-HNFCZKTMSA-N 0.000 description 2
- AGPKZVBTJJNPAG-WHFBIAKZSA-N L-isoleucine Chemical compound CC[C@H](C)[C@H](N)C(O)=O AGPKZVBTJJNPAG-WHFBIAKZSA-N 0.000 description 2
- ROHFNLRQFUQHCH-YFKPBYRVSA-N L-leucine Chemical compound CC(C)C[C@H](N)C(O)=O ROHFNLRQFUQHCH-YFKPBYRVSA-N 0.000 description 2
- COLNVLDHVKWLRT-QMMMGPOBSA-N L-phenylalanine Chemical compound OC(=O)[C@@H](N)CC1=CC=CC=C1 COLNVLDHVKWLRT-QMMMGPOBSA-N 0.000 description 2
- OGUUKPXUTHOIAV-SDDRHHMPSA-N Leu-Glu-Pro Chemical compound CC(C)C[C@@H](C(=O)N[C@@H](CCC(=O)O)C(=O)N1CCC[C@@H]1C(=O)O)N OGUUKPXUTHOIAV-SDDRHHMPSA-N 0.000 description 2
- APFJUBGRZGMQFF-QWRGUYRKSA-N Leu-Gly-Lys Chemical compound CC(C)C[C@H](N)C(=O)NCC(=O)N[C@H](C(O)=O)CCCCN APFJUBGRZGMQFF-QWRGUYRKSA-N 0.000 description 2
- QNBVTHNJGCOVFA-AVGNSLFASA-N Leu-Leu-Glu Chemical compound CC(C)C[C@H](N)C(=O)N[C@@H](CC(C)C)C(=O)N[C@H](C(O)=O)CCC(O)=O QNBVTHNJGCOVFA-AVGNSLFASA-N 0.000 description 2
- ROHFNLRQFUQHCH-UHFFFAOYSA-N Leucine Natural products CC(C)CC(N)C(O)=O ROHFNLRQFUQHCH-UHFFFAOYSA-N 0.000 description 2
- YVSHZSUKQHNDHD-KKUMJFAQSA-N Lys-Asn-Phe Chemical compound C1=CC=C(C=C1)C[C@@H](C(=O)O)NC(=O)[C@H](CC(=O)N)NC(=O)[C@H](CCCCN)N YVSHZSUKQHNDHD-KKUMJFAQSA-N 0.000 description 2
- ISHNZELVUVPCHY-ZETCQYMHSA-N Lys-Gly-Gly Chemical compound NCCCC[C@H](N)C(=O)NCC(=O)NCC(O)=O ISHNZELVUVPCHY-ZETCQYMHSA-N 0.000 description 2
- OVAOHZIOUBEQCJ-IHRRRGAJSA-N Lys-Leu-Arg Chemical compound [H]N[C@@H](CCCCN)C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](CCCNC(N)=N)C(O)=O OVAOHZIOUBEQCJ-IHRRRGAJSA-N 0.000 description 2
- WRODMZBHNNPRLN-SRVKXCTJSA-N Lys-Leu-Ser Chemical compound [H]N[C@@H](CCCCN)C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](CO)C(O)=O WRODMZBHNNPRLN-SRVKXCTJSA-N 0.000 description 2
- HVAUKHLDSDDROB-KKUMJFAQSA-N Lys-Lys-Leu Chemical compound [H]N[C@@H](CCCCN)C(=O)N[C@@H](CCCCN)C(=O)N[C@@H](CC(C)C)C(O)=O HVAUKHLDSDDROB-KKUMJFAQSA-N 0.000 description 2
- YDDDRTIPNTWGIG-SRVKXCTJSA-N Lys-Lys-Ser Chemical compound [H]N[C@@H](CCCCN)C(=O)N[C@@H](CCCCN)C(=O)N[C@@H](CO)C(O)=O YDDDRTIPNTWGIG-SRVKXCTJSA-N 0.000 description 2
- SBQDRNOLGSYHQA-YUMQZZPRSA-N Lys-Ser-Gly Chemical compound [H]N[C@@H](CCCCN)C(=O)N[C@@H](CO)C(=O)NCC(O)=O SBQDRNOLGSYHQA-YUMQZZPRSA-N 0.000 description 2
- WZVSHTFTCYOFPL-GARJFASQSA-N Lys-Ser-Pro Chemical compound C1C[C@@H](N(C1)C(=O)[C@H](CO)NC(=O)[C@H](CCCCN)N)C(=O)O WZVSHTFTCYOFPL-GARJFASQSA-N 0.000 description 2
- JKXVPNCSAMWUEJ-GUBZILKMSA-N Met-Met-Asp Chemical compound [H]N[C@@H](CCSC)C(=O)N[C@@H](CCSC)C(=O)N[C@@H](CC(O)=O)C(O)=O JKXVPNCSAMWUEJ-GUBZILKMSA-N 0.000 description 2
- YBAFDPFAUTYYRW-UHFFFAOYSA-N N-L-alpha-glutamyl-L-leucine Natural products CC(C)CC(C(O)=O)NC(=O)C(N)CCC(O)=O YBAFDPFAUTYYRW-UHFFFAOYSA-N 0.000 description 2
- OVRNDRQMDRJTHS-FMDGEEDCSA-N N-acetyl-beta-D-glucosamine Chemical compound CC(=O)N[C@H]1[C@H](O)O[C@H](CO)[C@@H](O)[C@@H]1O OVRNDRQMDRJTHS-FMDGEEDCSA-N 0.000 description 2
- FDJKUWYYUZCUJX-UHFFFAOYSA-N N-glycolyl-beta-neuraminic acid Natural products OCC(O)C(O)C1OC(O)(C(O)=O)CC(O)C1NC(=O)CO FDJKUWYYUZCUJX-UHFFFAOYSA-N 0.000 description 2
- FDJKUWYYUZCUJX-KVNVFURPSA-N N-glycolylneuraminic acid Chemical group OC[C@H](O)[C@H](O)[C@@H]1O[C@](O)(C(O)=O)C[C@H](O)[C@H]1NC(=O)CO FDJKUWYYUZCUJX-KVNVFURPSA-N 0.000 description 2
- LSXGADJXBDFXQU-DLOVCJGASA-N Phe-Ala-Asp Chemical compound OC(=O)C[C@@H](C(O)=O)NC(=O)[C@H](C)NC(=O)[C@@H](N)CC1=CC=CC=C1 LSXGADJXBDFXQU-DLOVCJGASA-N 0.000 description 2
- QPQDWBAJWOGAMJ-IHPCNDPISA-N Phe-Asp-Trp Chemical compound C([C@H](N)C(=O)N[C@@H](CC(O)=O)C(=O)N[C@@H](CC=1C2=CC=CC=C2NC=1)C(O)=O)C1=CC=CC=C1 QPQDWBAJWOGAMJ-IHPCNDPISA-N 0.000 description 2
- 229920002562 Polyethylene Glycol 3350 Polymers 0.000 description 2
- UIMCLYYSUCIUJM-UWVGGRQHSA-N Pro-Gly-Lys Chemical compound NCCCC[C@@H](C(O)=O)NC(=O)CNC(=O)[C@@H]1CCCN1 UIMCLYYSUCIUJM-UWVGGRQHSA-N 0.000 description 2
- STASJMBVVHNWCG-IHRRRGAJSA-N Pro-His-Leu Chemical compound C([C@@H](C(=O)N[C@@H](CC(C)C)C([O-])=O)NC(=O)[C@H]1[NH2+]CCC1)C1=CN=CN1 STASJMBVVHNWCG-IHRRRGAJSA-N 0.000 description 2
- SVXXJYJCRNKDDE-AVGNSLFASA-N Pro-Pro-His Chemical compound C([C@@H](C(=O)O)NC(=O)[C@H]1N(CCC1)C(=O)[C@H]1NCCC1)C1=CN=CN1 SVXXJYJCRNKDDE-AVGNSLFASA-N 0.000 description 2
- LZHHZYDPMZEMRX-STQMWFEESA-N Pro-Tyr-Gly Chemical compound [H]N1CCC[C@H]1C(=O)N[C@@H](CC1=CC=C(O)C=C1)C(=O)NCC(O)=O LZHHZYDPMZEMRX-STQMWFEESA-N 0.000 description 2
- 108010076504 Protein Sorting Signals Proteins 0.000 description 2
- 108020004511 Recombinant DNA Proteins 0.000 description 2
- FIXILCYTSAUERA-FXQIFTODSA-N Ser-Ala-Arg Chemical compound [H]N[C@@H](CO)C(=O)N[C@@H](C)C(=O)N[C@@H](CCCNC(N)=N)C(O)=O FIXILCYTSAUERA-FXQIFTODSA-N 0.000 description 2
- RZEQTVHJZCIUBT-WDSKDSINSA-N Ser-Arg Chemical compound OC[C@H](N)C(=O)N[C@H](C(O)=O)CCCNC(N)=N RZEQTVHJZCIUBT-WDSKDSINSA-N 0.000 description 2
- HBOABDXGTMMDSE-GUBZILKMSA-N Ser-Arg-Val Chemical compound [H]N[C@@H](CO)C(=O)N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](C(C)C)C(O)=O HBOABDXGTMMDSE-GUBZILKMSA-N 0.000 description 2
- KMWFXJCGRXBQAC-CIUDSAMLSA-N Ser-Cys-Lys Chemical compound C(CCN)C[C@@H](C(=O)O)NC(=O)[C@H](CS)NC(=O)[C@H](CO)N KMWFXJCGRXBQAC-CIUDSAMLSA-N 0.000 description 2
- NIOYDASGXWLHEZ-CIUDSAMLSA-N Ser-Met-Glu Chemical compound [H]N[C@@H](CO)C(=O)N[C@@H](CCSC)C(=O)N[C@@H](CCC(O)=O)C(O)=O NIOYDASGXWLHEZ-CIUDSAMLSA-N 0.000 description 2
- BEBVVQPDSHHWQL-NRPADANISA-N Ser-Val-Glu Chemical compound [H]N[C@@H](CO)C(=O)N[C@@H](C(C)C)C(=O)N[C@@H](CCC(O)=O)C(O)=O BEBVVQPDSHHWQL-NRPADANISA-N 0.000 description 2
- YEDSOSIKVUMIJE-DCAQKATOSA-N Ser-Val-Leu Chemical compound [H]N[C@@H](CO)C(=O)N[C@@H](C(C)C)C(=O)N[C@@H](CC(C)C)C(O)=O YEDSOSIKVUMIJE-DCAQKATOSA-N 0.000 description 2
- MTCFGRXMJLQNBG-UHFFFAOYSA-N Serine Natural products OCC(N)C(O)=O MTCFGRXMJLQNBG-UHFFFAOYSA-N 0.000 description 2
- PXIPVTKHYLBLMZ-UHFFFAOYSA-N Sodium azide Chemical compound [Na+].[N-]=[N+]=[N-] PXIPVTKHYLBLMZ-UHFFFAOYSA-N 0.000 description 2
- ZUXQFMVPAYGPFJ-JXUBOQSCSA-N Thr-Ala-Lys Chemical compound C[C@@H](O)[C@H](N)C(=O)N[C@@H](C)C(=O)N[C@H](C(O)=O)CCCCN ZUXQFMVPAYGPFJ-JXUBOQSCSA-N 0.000 description 2
- RFKVQLIXNVEOMB-WEDXCCLWSA-N Thr-Leu-Gly Chemical compound C[C@H]([C@@H](C(=O)N[C@@H](CC(C)C)C(=O)NCC(=O)O)N)O RFKVQLIXNVEOMB-WEDXCCLWSA-N 0.000 description 2
- NCXVJIQMWSGRHY-KXNHARMFSA-N Thr-Leu-Pro Chemical compound C[C@H]([C@@H](C(=O)N[C@@H](CC(C)C)C(=O)N1CCC[C@@H]1C(=O)O)N)O NCXVJIQMWSGRHY-KXNHARMFSA-N 0.000 description 2
- GUHLYMZJVXUIPO-RCWTZXSCSA-N Thr-Met-Val Chemical compound [H]N[C@@H]([C@@H](C)O)C(=O)N[C@@H](CCSC)C(=O)N[C@@H](C(C)C)C(O)=O GUHLYMZJVXUIPO-RCWTZXSCSA-N 0.000 description 2
- IEZVHOULSUULHD-XGEHTFHBSA-N Thr-Ser-Val Chemical compound [H]N[C@@H]([C@@H](C)O)C(=O)N[C@@H](CO)C(=O)N[C@@H](C(C)C)C(O)=O IEZVHOULSUULHD-XGEHTFHBSA-N 0.000 description 2
- SPIFGZFZMVLPHN-UNQGMJICSA-N Thr-Val-Phe Chemical compound [H]N[C@@H]([C@@H](C)O)C(=O)N[C@@H](C(C)C)C(=O)N[C@@H](CC1=CC=CC=C1)C(O)=O SPIFGZFZMVLPHN-UNQGMJICSA-N 0.000 description 2
- IQFYYKKMVGJFEH-XLPZGREQSA-N Thymidine Chemical compound O=C1NC(=O)C(C)=CN1[C@@H]1O[C@H](CO)[C@@H](O)C1 IQFYYKKMVGJFEH-XLPZGREQSA-N 0.000 description 2
- ADBFWLXCCKIXBQ-XIRDDKMYSA-N Trp-Asn-Leu Chemical compound CC(C)C[C@@H](C(=O)O)NC(=O)[C@H](CC(=O)N)NC(=O)[C@H](CC1=CNC2=CC=CC=C21)N ADBFWLXCCKIXBQ-XIRDDKMYSA-N 0.000 description 2
- NESIQDDPEFTWAH-BPUTZDHNSA-N Trp-Met-Asp Chemical compound [H]N[C@@H](CC1=CNC2=C1C=CC=C2)C(=O)N[C@@H](CCSC)C(=O)N[C@@H](CC(O)=O)C(O)=O NESIQDDPEFTWAH-BPUTZDHNSA-N 0.000 description 2
- YRBHLWWGSSQICE-IHRRRGAJSA-N Tyr-Asp-Met Chemical compound [H]N[C@@H](CC1=CC=C(O)C=C1)C(=O)N[C@@H](CC(O)=O)C(=O)N[C@@H](CCSC)C(O)=O YRBHLWWGSSQICE-IHRRRGAJSA-N 0.000 description 2
- IWRMTNJCCMEBEX-AVGNSLFASA-N Tyr-Glu-Cys Chemical compound C1=CC(=CC=C1C[C@@H](C(=O)N[C@@H](CCC(=O)O)C(=O)N[C@@H](CS)C(=O)O)N)O IWRMTNJCCMEBEX-AVGNSLFASA-N 0.000 description 2
- LRHBBGDMBLFYGL-FHWLQOOXSA-N Tyr-Phe-Glu Chemical compound C([C@H](N)C(=O)N[C@@H](CC=1C=CC=CC=1)C(=O)N[C@@H](CCC(O)=O)C(O)=O)C1=CC=C(O)C=C1 LRHBBGDMBLFYGL-FHWLQOOXSA-N 0.000 description 2
- QKXAEWMHAAVVGS-KKUMJFAQSA-N Tyr-Pro-Glu Chemical compound N[C@@H](Cc1ccc(O)cc1)C(=O)N1CCC[C@H]1C(=O)N[C@@H](CCC(O)=O)C(O)=O QKXAEWMHAAVVGS-KKUMJFAQSA-N 0.000 description 2
- MWUYSCVVPVITMW-IGNZVWTISA-N Tyr-Tyr-Ala Chemical compound C([C@@H](C(=O)N[C@@H](C)C(O)=O)NC(=O)[C@@H](N)CC=1C=CC(O)=CC=1)C1=CC=C(O)C=C1 MWUYSCVVPVITMW-IGNZVWTISA-N 0.000 description 2
- ISAKRJDGNUQOIC-UHFFFAOYSA-N Uracil Chemical compound O=C1C=CNC(=O)N1 ISAKRJDGNUQOIC-UHFFFAOYSA-N 0.000 description 2
- DRTQHJPVMGBUCF-XVFCMESISA-N Uridine Chemical compound O[C@@H]1[C@H](O)[C@@H](CO)O[C@H]1N1C(=O)NC(=O)C=C1 DRTQHJPVMGBUCF-XVFCMESISA-N 0.000 description 2
- XCCTYIAWTASOJW-XVFCMESISA-N Uridine-5'-Diphosphate Chemical compound O[C@@H]1[C@H](O)[C@@H](COP(O)(=O)OP(O)(O)=O)O[C@H]1N1C(=O)NC(=O)C=C1 XCCTYIAWTASOJW-XVFCMESISA-N 0.000 description 2
- COYSIHFOCOMGCF-UHFFFAOYSA-N Val-Arg-Gly Natural products CC(C)C(N)C(=O)NC(C(=O)NCC(O)=O)CCCN=C(N)N COYSIHFOCOMGCF-UHFFFAOYSA-N 0.000 description 2
- CWOSXNKDOACNJN-BZSNNMDCSA-N Val-Arg-Trp Chemical compound CC(C)[C@@H](C(=O)N[C@@H](CCCN=C(N)N)C(=O)N[C@@H](CC1=CNC2=CC=CC=C21)C(=O)O)N CWOSXNKDOACNJN-BZSNNMDCSA-N 0.000 description 2
- XXROXFHCMVXETG-UWVGGRQHSA-N Val-Gly-Val Chemical compound CC(C)[C@H](N)C(=O)NCC(=O)N[C@@H](C(C)C)C(O)=O XXROXFHCMVXETG-UWVGGRQHSA-N 0.000 description 2
- AEMPCGRFEZTWIF-IHRRRGAJSA-N Val-Leu-Lys Chemical compound CC(C)[C@H](N)C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](CCCCN)C(O)=O AEMPCGRFEZTWIF-IHRRRGAJSA-N 0.000 description 2
- SSYBNWFXCFNRFN-GUBZILKMSA-N Val-Pro-Ser Chemical compound CC(C)[C@H](N)C(=O)N1CCC[C@H]1C(=O)N[C@@H](CO)C(O)=O SSYBNWFXCFNRFN-GUBZILKMSA-N 0.000 description 2
- KZSNJWFQEVHDMF-UHFFFAOYSA-N Valine Natural products CC(C)C(N)C(O)=O KZSNJWFQEVHDMF-UHFFFAOYSA-N 0.000 description 2
- OIRDTQYFTABQOQ-KQYNXXCUSA-N adenosine Chemical compound C1=NC=2C(N)=NC=NC=2N1[C@@H]1O[C@H](CO)[C@@H](O)[C@H]1O OIRDTQYFTABQOQ-KQYNXXCUSA-N 0.000 description 2
- 235000004279 alanine Nutrition 0.000 description 2
- 108010008685 alanyl-glutamyl-aspartic acid Proteins 0.000 description 2
- 238000004458 analytical method Methods 0.000 description 2
- 239000003242 anti bacterial agent Substances 0.000 description 2
- 230000000692 anti-sense effect Effects 0.000 description 2
- 235000003704 aspartic acid Nutrition 0.000 description 2
- 108010038633 aspartylglutamate Proteins 0.000 description 2
- 108010047857 aspartylglycine Proteins 0.000 description 2
- OQFSQFPPLPISGP-UHFFFAOYSA-N beta-carboxyaspartic acid Natural products OC(=O)C(N)C(C(O)=O)C(O)=O OQFSQFPPLPISGP-UHFFFAOYSA-N 0.000 description 2
- 230000003115 biocidal effect Effects 0.000 description 2
- 230000037396 body weight Effects 0.000 description 2
- 210000004899 c-terminal region Anatomy 0.000 description 2
- 125000003178 carboxy group Chemical group [H]OC(*)=O 0.000 description 2
- 230000001413 cellular effect Effects 0.000 description 2
- 239000000356 contaminant Substances 0.000 description 2
- 230000008878 coupling Effects 0.000 description 2
- 238000010168 coupling process Methods 0.000 description 2
- 238000005859 coupling reaction Methods 0.000 description 2
- 235000018417 cysteine Nutrition 0.000 description 2
- XUJNEKJLAYXESH-UHFFFAOYSA-N cysteine Natural products SCC(N)C(O)=O XUJNEKJLAYXESH-UHFFFAOYSA-N 0.000 description 2
- OPTASPLRGRRNAP-UHFFFAOYSA-N cytosine Chemical compound NC=1C=CNC(=O)N=1 OPTASPLRGRRNAP-UHFFFAOYSA-N 0.000 description 2
- 201000010099 disease Diseases 0.000 description 2
- 208000037265 diseases, disorders, signs and symptoms Diseases 0.000 description 2
- 238000005516 engineering process Methods 0.000 description 2
- 239000006167 equilibration buffer Substances 0.000 description 2
- ZDXPYRJPNDTMRX-UHFFFAOYSA-N glutamine Natural products OC(=O)C(N)CCC(N)=O ZDXPYRJPNDTMRX-UHFFFAOYSA-N 0.000 description 2
- 108010067216 glycyl-glycyl-glycine Proteins 0.000 description 2
- 108010081551 glycylphenylalanine Proteins 0.000 description 2
- 229960004198 guanidine Drugs 0.000 description 2
- PJJJBBJSCAKJQF-UHFFFAOYSA-N guanidinium chloride Chemical compound [Cl-].NC(N)=[NH2+] PJJJBBJSCAKJQF-UHFFFAOYSA-N 0.000 description 2
- UYTPUPDQBNUYGX-UHFFFAOYSA-N guanine Chemical compound O=C1NC(N)=NC2=C1N=CN2 UYTPUPDQBNUYGX-UHFFFAOYSA-N 0.000 description 2
- 239000012145 high-salt buffer Substances 0.000 description 2
- 108010028295 histidylhistidine Proteins 0.000 description 2
- 230000002209 hydrophobic effect Effects 0.000 description 2
- 238000001727 in vivo Methods 0.000 description 2
- 108010034529 leucyl-lysine Proteins 0.000 description 2
- 108010073472 leucyl-prolyl-proline Proteins 0.000 description 2
- 239000002502 liposome Substances 0.000 description 2
- 229920002521 macromolecule Polymers 0.000 description 2
- 238000013507 mapping Methods 0.000 description 2
- 239000011159 matrix material Substances 0.000 description 2
- 238000001840 matrix-assisted laser desorption--ionisation time-of-flight mass spectrometry Methods 0.000 description 2
- 238000002156 mixing Methods 0.000 description 2
- 230000035772 mutation Effects 0.000 description 2
- 229940060155 neuac Drugs 0.000 description 2
- CERZMXAJYMMUDR-UHFFFAOYSA-N neuraminic acid Natural products NC1C(O)CC(O)(C(O)=O)OC1C(O)C(O)CO CERZMXAJYMMUDR-UHFFFAOYSA-N 0.000 description 2
- COLNVLDHVKWLRT-UHFFFAOYSA-N phenylalanine Natural products OC(=O)C(N)CC1=CC=CC=C1 COLNVLDHVKWLRT-UHFFFAOYSA-N 0.000 description 2
- 108010012581 phenylalanylglutamate Proteins 0.000 description 2
- 108010073101 phenylalanylleucine Proteins 0.000 description 2
- 230000008488 polyadenylation Effects 0.000 description 2
- 108010031719 prolyl-serine Proteins 0.000 description 2
- 108010090894 prolylleucine Proteins 0.000 description 2
- 230000001902 propagating effect Effects 0.000 description 2
- 108020001580 protein domains Proteins 0.000 description 2
- 230000017854 proteolysis Effects 0.000 description 2
- 239000011535 reaction buffer Substances 0.000 description 2
- 230000003362 replicative effect Effects 0.000 description 2
- 102220172123 rs886048669 Human genes 0.000 description 2
- 150000003839 salts Chemical class 0.000 description 2
- 238000012216 screening Methods 0.000 description 2
- 108010026333 seryl-proline Proteins 0.000 description 2
- 230000003381 solubilizing effect Effects 0.000 description 2
- 238000001228 spectrum Methods 0.000 description 2
- 108010005652 splenotritin Proteins 0.000 description 2
- 150000008163 sugars Chemical class 0.000 description 2
- 238000012360 testing method Methods 0.000 description 2
- 125000000341 threoninyl group Chemical group [H]OC([H])(C([H])([H])[H])C([H])(N([H])[H])C(*)=O 0.000 description 2
- RWQNBRDOKXIBIV-UHFFFAOYSA-N thymine Chemical compound CC1=CNC(=O)NC1=O RWQNBRDOKXIBIV-UHFFFAOYSA-N 0.000 description 2
- 238000013519 translation Methods 0.000 description 2
- 108010084932 tryptophyl-proline Proteins 0.000 description 2
- 108010051110 tyrosyl-lysine Proteins 0.000 description 2
- 108010020532 tyrosyl-proline Proteins 0.000 description 2
- 238000011144 upstream manufacturing Methods 0.000 description 2
- UHDGCWIWMRVCDJ-UHFFFAOYSA-N 1-beta-D-Xylofuranosyl-NH-Cytosine Natural products O=C1N=C(N)C=CN1C1C(O)C(O)C(CO)O1 UHDGCWIWMRVCDJ-UHFFFAOYSA-N 0.000 description 1
- MSWZFWKMSRAUBD-IVMDWMLBSA-N 2-amino-2-deoxy-D-glucopyranose Chemical compound N[C@H]1C(O)O[C@H](CO)[C@@H](O)[C@@H]1O MSWZFWKMSRAUBD-IVMDWMLBSA-N 0.000 description 1
- FWMNVWWHGCHHJJ-SKKKGAJSSA-N 4-amino-1-[(2r)-6-amino-2-[[(2r)-2-[[(2r)-2-[[(2r)-2-amino-3-phenylpropanoyl]amino]-3-phenylpropanoyl]amino]-4-methylpentanoyl]amino]hexanoyl]piperidine-4-carboxylic acid Chemical compound C([C@H](C(=O)N[C@H](CC(C)C)C(=O)N[C@H](CCCCN)C(=O)N1CCC(N)(CC1)C(O)=O)NC(=O)[C@H](N)CC=1C=CC=CC=1)C1=CC=CC=C1 FWMNVWWHGCHHJJ-SKKKGAJSSA-N 0.000 description 1
- 102000007469 Actins Human genes 0.000 description 1
- 108010085238 Actins Proteins 0.000 description 1
- 229930024421 Adenine Natural products 0.000 description 1
- GFFGJBXGBJISGV-UHFFFAOYSA-N Adenine Chemical compound NC1=NC=NC2=C1N=CN2 GFFGJBXGBJISGV-UHFFFAOYSA-N 0.000 description 1
- NHCPCLJZRSIDHS-ZLUOBGJFSA-N Ala-Asp-Ala Chemical compound [H]N[C@@H](C)C(=O)N[C@@H](CC(O)=O)C(=O)N[C@@H](C)C(O)=O NHCPCLJZRSIDHS-ZLUOBGJFSA-N 0.000 description 1
- KIUYPHAMDKDICO-WHFBIAKZSA-N Ala-Asp-Gly Chemical compound C[C@H](N)C(=O)N[C@@H](CC(O)=O)C(=O)NCC(O)=O KIUYPHAMDKDICO-WHFBIAKZSA-N 0.000 description 1
- ZIWWTZWAKYBUOB-CIUDSAMLSA-N Ala-Asp-Leu Chemical compound [H]N[C@@H](C)C(=O)N[C@@H](CC(O)=O)C(=O)N[C@@H](CC(C)C)C(O)=O ZIWWTZWAKYBUOB-CIUDSAMLSA-N 0.000 description 1
- OMMDTNGURYRDAC-NRPADANISA-N Ala-Glu-Val Chemical compound [H]N[C@@H](C)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](C(C)C)C(O)=O OMMDTNGURYRDAC-NRPADANISA-N 0.000 description 1
- NINQYGGNRIBFSC-CIUDSAMLSA-N Ala-Lys-Ser Chemical compound NCCCC[C@H](NC(=O)[C@@H](N)C)C(=O)N[C@@H](CO)C(O)=O NINQYGGNRIBFSC-CIUDSAMLSA-N 0.000 description 1
- ZBLQIYPCUWZSRZ-QEJZJMRPSA-N Ala-Phe-Leu Chemical compound CC(C)C[C@@H](C(O)=O)NC(=O)[C@@H](NC(=O)[C@H](C)N)CC1=CC=CC=C1 ZBLQIYPCUWZSRZ-QEJZJMRPSA-N 0.000 description 1
- MUGAESARFRGOTQ-IGNZVWTISA-N Ala-Tyr-Tyr Chemical compound C[C@@H](C(=O)N[C@@H](CC1=CC=C(C=C1)O)C(=O)N[C@@H](CC2=CC=C(C=C2)O)C(=O)O)N MUGAESARFRGOTQ-IGNZVWTISA-N 0.000 description 1
- 108020005098 Anticodon Proteins 0.000 description 1
- ZTKHZAXGTFXUDD-VEVYYDQMSA-N Arg-Asn-Thr Chemical compound [H]N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CC(N)=O)C(=O)N[C@@H]([C@@H](C)O)C(O)=O ZTKHZAXGTFXUDD-VEVYYDQMSA-N 0.000 description 1
- PBSOQGZLPFVXPU-YUMQZZPRSA-N Arg-Glu-Gly Chemical compound [H]N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CCC(O)=O)C(=O)NCC(O)=O PBSOQGZLPFVXPU-YUMQZZPRSA-N 0.000 description 1
- BTJVOUQWFXABOI-IHRRRGAJSA-N Arg-Lys-Lys Chemical compound NCCCC[C@@H](C(O)=O)NC(=O)[C@H](CCCCN)NC(=O)[C@@H](N)CCCNC(N)=N BTJVOUQWFXABOI-IHRRRGAJSA-N 0.000 description 1
- OISWSORSLQOGFV-AVGNSLFASA-N Arg-Met-Leu Chemical compound CC(C)C[C@@H](C(O)=O)NC(=O)[C@H](CCSC)NC(=O)[C@@H](N)CCCN=C(N)N OISWSORSLQOGFV-AVGNSLFASA-N 0.000 description 1
- XRNXPIGJPQHCPC-RCWTZXSCSA-N Arg-Thr-Val Chemical compound CC(C)[C@H](NC(=O)[C@@H](NC(=O)[C@@H](N)CCCNC(N)=N)[C@@H](C)O)C(O)=O XRNXPIGJPQHCPC-RCWTZXSCSA-N 0.000 description 1
- QTAIIXQCOPUNBQ-QXEWZRGKSA-N Arg-Val-Asp Chemical compound [H]N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](C(C)C)C(=O)N[C@@H](CC(O)=O)C(O)=O QTAIIXQCOPUNBQ-QXEWZRGKSA-N 0.000 description 1
- HZPSDHRYYIORKR-WHFBIAKZSA-N Asn-Ala-Gly Chemical compound OC(=O)CNC(=O)[C@H](C)NC(=O)[C@@H](N)CC(N)=O HZPSDHRYYIORKR-WHFBIAKZSA-N 0.000 description 1
- WVCJSDCHTUTONA-FXQIFTODSA-N Asn-Asp-Arg Chemical compound [H]N[C@@H](CC(N)=O)C(=O)N[C@@H](CC(O)=O)C(=O)N[C@@H](CCCNC(N)=N)C(O)=O WVCJSDCHTUTONA-FXQIFTODSA-N 0.000 description 1
- OGMDXNFGPOPZTK-GUBZILKMSA-N Asn-Glu-His Chemical compound C1=C(NC=N1)C[C@@H](C(=O)O)NC(=O)[C@H](CCC(=O)O)NC(=O)[C@H](CC(=O)N)N OGMDXNFGPOPZTK-GUBZILKMSA-N 0.000 description 1
- OPEPUCYIGFEGSW-WDSKDSINSA-N Asn-Gly-Glu Chemical compound [H]N[C@@H](CC(N)=O)C(=O)NCC(=O)N[C@@H](CCC(O)=O)C(O)=O OPEPUCYIGFEGSW-WDSKDSINSA-N 0.000 description 1
- UHGUKCOQUNPSKK-CIUDSAMLSA-N Asn-Leu-Cys Chemical compound CC(C)C[C@@H](C(=O)N[C@@H](CS)C(=O)O)NC(=O)[C@H](CC(=O)N)N UHGUKCOQUNPSKK-CIUDSAMLSA-N 0.000 description 1
- OMSMPWHEGLNQOD-UWVGGRQHSA-N Asn-Phe Chemical compound NC(=O)C[C@H](N)C(=O)N[C@H](C(O)=O)CC1=CC=CC=C1 OMSMPWHEGLNQOD-UWVGGRQHSA-N 0.000 description 1
- NYLBGYLHBDFRHL-VEVYYDQMSA-N Asp-Arg-Thr Chemical compound [H]N[C@@H](CC(O)=O)C(=O)N[C@@H](CCCNC(N)=N)C(=O)N[C@@H]([C@@H](C)O)C(O)=O NYLBGYLHBDFRHL-VEVYYDQMSA-N 0.000 description 1
- VZNOVQKGJQJOCS-SRVKXCTJSA-N Asp-Asp-Tyr Chemical compound [H]N[C@@H](CC(O)=O)C(=O)N[C@@H](CC(O)=O)C(=O)N[C@@H](CC1=CC=C(O)C=C1)C(O)=O VZNOVQKGJQJOCS-SRVKXCTJSA-N 0.000 description 1
- HSPSXROIMXIJQW-BQBZGAKWSA-N Asp-His Chemical compound OC(=O)C[C@H](N)C(=O)N[C@H](C(O)=O)CC1=CNC=N1 HSPSXROIMXIJQW-BQBZGAKWSA-N 0.000 description 1
- NBKLEMWHDLAUEM-CIUDSAMLSA-N Asp-Ser-His Chemical compound C1=C(NC=N1)C[C@@H](C(=O)O)NC(=O)[C@H](CO)NC(=O)[C@H](CC(=O)O)N NBKLEMWHDLAUEM-CIUDSAMLSA-N 0.000 description 1
- JSNWZMFSLIWAHS-HJGDQZAQSA-N Asp-Thr-Leu Chemical compound C[C@H]([C@@H](C(=O)N[C@@H](CC(C)C)C(=O)O)NC(=O)[C@H](CC(=O)O)N)O JSNWZMFSLIWAHS-HJGDQZAQSA-N 0.000 description 1
- WOKXEQLPBLLWHC-IHRRRGAJSA-N Asp-Tyr-Met Chemical compound CSCC[C@@H](C(O)=O)NC(=O)[C@@H](NC(=O)[C@@H](N)CC(O)=O)CC1=CC=C(O)C=C1 WOKXEQLPBLLWHC-IHRRRGAJSA-N 0.000 description 1
- GYNUXDMCDILYIQ-QRTARXTBSA-N Asp-Val-Trp Chemical compound CC(C)[C@@H](C(=O)N[C@@H](CC1=CNC2=CC=CC=C21)C(=O)O)NC(=O)[C@H](CC(=O)O)N GYNUXDMCDILYIQ-QRTARXTBSA-N 0.000 description 1
- 241000894006 Bacteria Species 0.000 description 1
- DWRXFEITVBNRMK-UHFFFAOYSA-N Beta-D-1-Arabinofuranosylthymine Natural products O=C1NC(=O)C(C)=CN1C1C(O)C(O)C(CO)O1 DWRXFEITVBNRMK-UHFFFAOYSA-N 0.000 description 1
- 102100021277 Beta-secretase 2 Human genes 0.000 description 1
- 101710150190 Beta-secretase 2 Proteins 0.000 description 1
- BTBUEUYNUDRHOZ-UHFFFAOYSA-N Borate Chemical compound [O-]B([O-])[O-] BTBUEUYNUDRHOZ-UHFFFAOYSA-N 0.000 description 1
- 239000002126 C01EB10 - Adenosine Substances 0.000 description 1
- 101100505161 Caenorhabditis elegans mel-32 gene Proteins 0.000 description 1
- 241000282693 Cercopithecidae Species 0.000 description 1
- VEXZGXHMUGYJMC-UHFFFAOYSA-M Chloride anion Chemical group [Cl-] VEXZGXHMUGYJMC-UHFFFAOYSA-M 0.000 description 1
- 101100359525 Chlorobium chlorochromatii (strain CaD3) rpmI gene Proteins 0.000 description 1
- 102100031263 Chondroitin sulfate N-acetylgalactosaminyltransferase 2 Human genes 0.000 description 1
- 108091062157 Cis-regulatory element Proteins 0.000 description 1
- 241000699802 Cricetulus griseus Species 0.000 description 1
- MIKUYHXYGGJMLM-GIMIYPNGSA-N Crotonoside Natural products C1=NC2=C(N)NC(=O)N=C2N1[C@H]1O[C@@H](CO)[C@H](O)[C@@H]1O MIKUYHXYGGJMLM-GIMIYPNGSA-N 0.000 description 1
- CLDCTNHPILWQCW-CIUDSAMLSA-N Cys-Arg-Glu Chemical compound C(C[C@@H](C(=O)N[C@@H](CCC(=O)O)C(=O)O)NC(=O)[C@H](CS)N)CN=C(N)N CLDCTNHPILWQCW-CIUDSAMLSA-N 0.000 description 1
- BCSYBBMFGLHCOA-ACZMJKKPSA-N Cys-Glu-Cys Chemical compound SC[C@H](N)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CS)C(O)=O BCSYBBMFGLHCOA-ACZMJKKPSA-N 0.000 description 1
- SKSJPIBFNFPTJB-NKWVEPMBSA-N Cys-Gly-Pro Chemical compound C1C[C@@H](N(C1)C(=O)CNC(=O)[C@H](CS)N)C(=O)O SKSJPIBFNFPTJB-NKWVEPMBSA-N 0.000 description 1
- VNXXMHTZQGGDSG-CIUDSAMLSA-N Cys-His-Asn Chemical compound [H]N[C@@H](CS)C(=O)N[C@@H](CC1=CNC=N1)C(=O)N[C@@H](CC(N)=O)C(O)=O VNXXMHTZQGGDSG-CIUDSAMLSA-N 0.000 description 1
- UHDGCWIWMRVCDJ-PSQAKQOGSA-N Cytidine Natural products O=C1N=C(N)C=CN1[C@@H]1[C@@H](O)[C@@H](O)[C@H](CO)O1 UHDGCWIWMRVCDJ-PSQAKQOGSA-N 0.000 description 1
- 150000008574 D-amino acids Chemical class 0.000 description 1
- LEVWYRKDKASIDU-QWWZWVQMSA-N D-cystine Chemical compound OC(=O)[C@H](N)CSSC[C@@H](N)C(O)=O LEVWYRKDKASIDU-QWWZWVQMSA-N 0.000 description 1
- 229930182847 D-glutamic acid Natural products 0.000 description 1
- NYHBQMYGNKIUIF-UHFFFAOYSA-N D-guanosine Natural products C1=2NC(N)=NC(=O)C=2N=CN1C1OC(CO)C(O)C1O NYHBQMYGNKIUIF-UHFFFAOYSA-N 0.000 description 1
- 102000016928 DNA-directed DNA polymerase Human genes 0.000 description 1
- 108010014303 DNA-directed DNA polymerase Proteins 0.000 description 1
- 241000702421 Dependoparvovirus Species 0.000 description 1
- 229920002307 Dextran Polymers 0.000 description 1
- BWGNESOTFCXPMA-UHFFFAOYSA-N Dihydrogen disulfide Chemical compound SS BWGNESOTFCXPMA-UHFFFAOYSA-N 0.000 description 1
- YQYJSBFKSSDGFO-UHFFFAOYSA-N Epihygromycin Natural products OC1C(O)C(C(=O)C)OC1OC(C(=C1)O)=CC=C1C=C(C)C(=O)NC1C(O)C(O)C2OCOC2C1O YQYJSBFKSSDGFO-UHFFFAOYSA-N 0.000 description 1
- 241000206602 Eukaryota Species 0.000 description 1
- WSFSSNUMVMOOMR-UHFFFAOYSA-N Formaldehyde Chemical compound O=C WSFSSNUMVMOOMR-UHFFFAOYSA-N 0.000 description 1
- 210000000712 G cell Anatomy 0.000 description 1
- PNAOVYHADQRJQU-GUBZILKMSA-N Glu-Cys-His Chemical compound C1=C(NC=N1)C[C@@H](C(=O)O)NC(=O)[C@H](CS)NC(=O)[C@H](CCC(=O)O)N PNAOVYHADQRJQU-GUBZILKMSA-N 0.000 description 1
- BBBXWRGITSUJPB-YUMQZZPRSA-N Glu-Lys Chemical compound NCCCC[C@@H](C(O)=O)NC(=O)[C@@H](N)CCC(O)=O BBBXWRGITSUJPB-YUMQZZPRSA-N 0.000 description 1
- OQXDUSZKISQQSS-GUBZILKMSA-N Glu-Lys-Ala Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N[C@@H](CCCCN)C(=O)N[C@@H](C)C(O)=O OQXDUSZKISQQSS-GUBZILKMSA-N 0.000 description 1
- SYWCGQOIIARSIX-SRVKXCTJSA-N Glu-Pro-Leu Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N1CCC[C@H]1C(=O)N[C@@H](CC(C)C)C(O)=O SYWCGQOIIARSIX-SRVKXCTJSA-N 0.000 description 1
- KCCNSVHJSMMGFS-NRPADANISA-N Glu-Val-Cys Chemical compound CC(C)[C@@H](C(=O)N[C@@H](CS)C(=O)O)NC(=O)[C@H](CCC(=O)O)N KCCNSVHJSMMGFS-NRPADANISA-N 0.000 description 1
- 108010092364 Glucuronosyltransferase Proteins 0.000 description 1
- 102000016354 Glucuronosyltransferase Human genes 0.000 description 1
- 108010024636 Glutathione Proteins 0.000 description 1
- VSVZIEVNUYDAFR-YUMQZZPRSA-N Gly-Ala-Leu Chemical compound CC(C)C[C@@H](C(O)=O)NC(=O)[C@H](C)NC(=O)CN VSVZIEVNUYDAFR-YUMQZZPRSA-N 0.000 description 1
- LJPIRKICOISLKN-WHFBIAKZSA-N Gly-Ala-Ser Chemical compound NCC(=O)N[C@@H](C)C(=O)N[C@@H](CO)C(O)=O LJPIRKICOISLKN-WHFBIAKZSA-N 0.000 description 1
- XCLCVBYNGXEVDU-WHFBIAKZSA-N Gly-Asn-Ser Chemical compound NCC(=O)N[C@@H](CC(N)=O)C(=O)N[C@@H](CO)C(O)=O XCLCVBYNGXEVDU-WHFBIAKZSA-N 0.000 description 1
- GDOZQTNZPCUARW-YFKPBYRVSA-N Gly-Gly-Glu Chemical compound NCC(=O)NCC(=O)N[C@H](C(O)=O)CCC(O)=O GDOZQTNZPCUARW-YFKPBYRVSA-N 0.000 description 1
- HPAIKDPJURGQLN-KBPBESRZSA-N Gly-His-Phe Chemical compound C([C@H](NC(=O)CN)C(=O)N[C@@H](CC=1C=CC=CC=1)C(O)=O)C1=CNC=N1 HPAIKDPJURGQLN-KBPBESRZSA-N 0.000 description 1
- GAFKBWKVXNERFA-QWRGUYRKSA-N Gly-Phe-Asp Chemical compound OC(=O)C[C@@H](C(O)=O)NC(=O)[C@@H](NC(=O)CN)CC1=CC=CC=C1 GAFKBWKVXNERFA-QWRGUYRKSA-N 0.000 description 1
- GGLIDLCEPDHEJO-BQBZGAKWSA-N Gly-Pro-Ala Chemical compound OC(=O)[C@H](C)NC(=O)[C@@H]1CCCN1C(=O)CN GGLIDLCEPDHEJO-BQBZGAKWSA-N 0.000 description 1
- WNGHUXFWEWTKAO-YUMQZZPRSA-N Gly-Ser-Leu Chemical compound CC(C)C[C@@H](C(O)=O)NC(=O)[C@H](CO)NC(=O)CN WNGHUXFWEWTKAO-YUMQZZPRSA-N 0.000 description 1
- 229930186217 Glycolipid Natural products 0.000 description 1
- KYMUEAZVLPRVAE-GUBZILKMSA-N His-Asn-Glu Chemical compound [H]N[C@@H](CC1=CNC=N1)C(=O)N[C@@H](CC(N)=O)C(=O)N[C@@H](CCC(O)=O)C(O)=O KYMUEAZVLPRVAE-GUBZILKMSA-N 0.000 description 1
- BZKDJRSZWLPJNI-SRVKXCTJSA-N His-His-Ser Chemical compound [H]N[C@@H](CC1=CNC=N1)C(=O)N[C@@H](CC1=CNC=N1)C(=O)N[C@@H](CO)C(O)=O BZKDJRSZWLPJNI-SRVKXCTJSA-N 0.000 description 1
- XDIVYNSPYBLSME-DCAQKATOSA-N His-Met-Asp Chemical compound CSCC[C@@H](C(=O)N[C@@H](CC(=O)O)C(=O)O)NC(=O)[C@H](CC1=CN=CN1)N XDIVYNSPYBLSME-DCAQKATOSA-N 0.000 description 1
- SAPLASXFNUYUFE-CQDKDKBSSA-N His-Phe-Ala Chemical compound C[C@@H](C(=O)O)NC(=O)[C@H](CC1=CC=CC=C1)NC(=O)[C@H](CC2=CN=CN2)N SAPLASXFNUYUFE-CQDKDKBSSA-N 0.000 description 1
- LNCFUHAPNTYMJB-IUCAKERBSA-N His-Pro Chemical compound C([C@H](N)C(=O)N1[C@@H](CCC1)C(O)=O)C1=CN=CN1 LNCFUHAPNTYMJB-IUCAKERBSA-N 0.000 description 1
- LNVILFYCPVOHPV-IHPCNDPISA-N His-Trp-Leu Chemical compound [H]N[C@@H](CC1=CNC=N1)C(=O)N[C@@H](CC1=CNC2=C1C=CC=C2)C(=O)N[C@@H](CC(C)C)C(O)=O LNVILFYCPVOHPV-IHPCNDPISA-N 0.000 description 1
- CGAMSLMBYJHMDY-ONGXEEELSA-N His-Val-Gly Chemical compound CC(C)[C@@H](C(=O)NCC(=O)O)NC(=O)[C@H](CC1=CN=CN1)N CGAMSLMBYJHMDY-ONGXEEELSA-N 0.000 description 1
- PMMYEEVYMWASQN-DMTCNVIQSA-N Hydroxyproline Chemical compound O[C@H]1CN[C@H](C(O)=O)C1 PMMYEEVYMWASQN-DMTCNVIQSA-N 0.000 description 1
- 108010065920 Insulin Lispro Proteins 0.000 description 1
- 102100039948 Interferon alpha-5 Human genes 0.000 description 1
- 101710106106 Interferon alpha-G Proteins 0.000 description 1
- 108010047761 Interferon-alpha Proteins 0.000 description 1
- 102000006992 Interferon-alpha Human genes 0.000 description 1
- 102000003996 Interferon-beta Human genes 0.000 description 1
- 108090000467 Interferon-beta Proteins 0.000 description 1
- 102000008070 Interferon-gamma Human genes 0.000 description 1
- 108010074328 Interferon-gamma Proteins 0.000 description 1
- 108091092195 Intron Proteins 0.000 description 1
- FADYJNXDPBKVCA-UHFFFAOYSA-N L-Phenylalanyl-L-lysin Natural products NCCCCC(C(O)=O)NC(=O)C(N)CC1=CC=CC=C1 FADYJNXDPBKVCA-UHFFFAOYSA-N 0.000 description 1
- QLROSWPKSBORFJ-BQBZGAKWSA-N L-Prolyl-L-glutamic acid Chemical compound OC(=O)CC[C@@H](C(O)=O)NC(=O)[C@@H]1CCCN1 QLROSWPKSBORFJ-BQBZGAKWSA-N 0.000 description 1
- 150000008575 L-amino acids Chemical class 0.000 description 1
- 229930064664 L-arginine Natural products 0.000 description 1
- 235000014852 L-arginine Nutrition 0.000 description 1
- ZDXPYRJPNDTMRX-VKHMYHEASA-N L-glutamine Chemical compound OC(=O)[C@@H](N)CCC(N)=O ZDXPYRJPNDTMRX-VKHMYHEASA-N 0.000 description 1
- KFKWRHQBZQICHA-STQMWFEESA-N L-leucyl-L-phenylalanine Natural products CC(C)C[C@H](N)C(=O)N[C@H](C(O)=O)CC1=CC=CC=C1 KFKWRHQBZQICHA-STQMWFEESA-N 0.000 description 1
- FFEARJCKVFRZRR-BYPYZUCNSA-N L-methionine Chemical compound CSCC[C@H](N)C(O)=O FFEARJCKVFRZRR-BYPYZUCNSA-N 0.000 description 1
- FBOZXECLQNJBKD-ZDUSSCGKSA-N L-methotrexate Chemical compound C=1N=C2N=C(N)N=C(N)C2=NC=1CN(C)C1=CC=C(C(=O)N[C@@H](CCC(O)=O)C(O)=O)C=C1 FBOZXECLQNJBKD-ZDUSSCGKSA-N 0.000 description 1
- KZSNJWFQEVHDMF-BYPYZUCNSA-N L-valine Chemical compound CC(C)[C@H](N)C(O)=O KZSNJWFQEVHDMF-BYPYZUCNSA-N 0.000 description 1
- 241000713666 Lentivirus Species 0.000 description 1
- YOZCKMXHBYKOMQ-IHRRRGAJSA-N Leu-Arg-Lys Chemical compound CC(C)C[C@@H](C(=O)N[C@@H](CCCN=C(N)N)C(=O)N[C@@H](CCCCN)C(=O)O)N YOZCKMXHBYKOMQ-IHRRRGAJSA-N 0.000 description 1
- STAVRDQLZOTNKJ-RHYQMDGZSA-N Leu-Arg-Thr Chemical compound [H]N[C@@H](CC(C)C)C(=O)N[C@@H](CCCNC(N)=N)C(=O)N[C@@H]([C@@H](C)O)C(O)=O STAVRDQLZOTNKJ-RHYQMDGZSA-N 0.000 description 1
- JKGHDYGZRDWHGA-SRVKXCTJSA-N Leu-Asn-Leu Chemical compound [H]N[C@@H](CC(C)C)C(=O)N[C@@H](CC(N)=O)C(=O)N[C@@H](CC(C)C)C(O)=O JKGHDYGZRDWHGA-SRVKXCTJSA-N 0.000 description 1
- CLVUXCBGKUECIT-HJGDQZAQSA-N Leu-Asp-Thr Chemical compound [H]N[C@@H](CC(C)C)C(=O)N[C@@H](CC(O)=O)C(=O)N[C@@H]([C@@H](C)O)C(O)=O CLVUXCBGKUECIT-HJGDQZAQSA-N 0.000 description 1
- YORLGJINWYYIMX-KKUMJFAQSA-N Leu-Cys-Phe Chemical compound [H]N[C@@H](CC(C)C)C(=O)N[C@@H](CS)C(=O)N[C@@H](CC1=CC=CC=C1)C(O)=O YORLGJINWYYIMX-KKUMJFAQSA-N 0.000 description 1
- QVFGXCVIXXBFHO-AVGNSLFASA-N Leu-Glu-Leu Chemical compound CC(C)C[C@H](N)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CC(C)C)C(O)=O QVFGXCVIXXBFHO-AVGNSLFASA-N 0.000 description 1
- LESXFEZIFXFIQR-LURJTMIESA-N Leu-Gly Chemical compound CC(C)C[C@H](N)C(=O)NCC(O)=O LESXFEZIFXFIQR-LURJTMIESA-N 0.000 description 1
- VBZOAGIPCULURB-QWRGUYRKSA-N Leu-Gly-His Chemical compound CC(C)C[C@@H](C(=O)NCC(=O)N[C@@H](CC1=CN=CN1)C(=O)O)N VBZOAGIPCULURB-QWRGUYRKSA-N 0.000 description 1
- KPYAOIVPJKPIOU-KKUMJFAQSA-N Leu-Lys-Lys Chemical compound CC(C)C[C@H](N)C(=O)N[C@@H](CCCCN)C(=O)N[C@@H](CCCCN)C(O)=O KPYAOIVPJKPIOU-KKUMJFAQSA-N 0.000 description 1
- PKKMDPNFGULLNQ-AVGNSLFASA-N Leu-Met-Arg Chemical compound CC(C)C[C@H](N)C(=O)N[C@@H](CCSC)C(=O)N[C@@H](CCCN=C(N)N)C(O)=O PKKMDPNFGULLNQ-AVGNSLFASA-N 0.000 description 1
- WMIOEVKKYIMVKI-DCAQKATOSA-N Leu-Pro-Ala Chemical compound [H]N[C@@H](CC(C)C)C(=O)N1CCC[C@H]1C(=O)N[C@@H](C)C(O)=O WMIOEVKKYIMVKI-DCAQKATOSA-N 0.000 description 1
- DPURXCQCHSQPAN-AVGNSLFASA-N Leu-Pro-Pro Chemical compound CC(C)C[C@H](N)C(=O)N1CCC[C@H]1C(=O)N1[C@H](C(O)=O)CCC1 DPURXCQCHSQPAN-AVGNSLFASA-N 0.000 description 1
- XGDCYUQSFDQISZ-BQBZGAKWSA-N Leu-Ser Chemical compound CC(C)C[C@H](N)C(=O)N[C@@H](CO)C(O)=O XGDCYUQSFDQISZ-BQBZGAKWSA-N 0.000 description 1
- KIZIOFNVSOSKJI-CIUDSAMLSA-N Leu-Ser-Cys Chemical compound CC(C)C[C@@H](C(=O)N[C@@H](CO)C(=O)N[C@@H](CS)C(=O)O)N KIZIOFNVSOSKJI-CIUDSAMLSA-N 0.000 description 1
- SVBJIZVVYJYGLA-DCAQKATOSA-N Leu-Ser-Val Chemical compound [H]N[C@@H](CC(C)C)C(=O)N[C@@H](CO)C(=O)N[C@@H](C(C)C)C(O)=O SVBJIZVVYJYGLA-DCAQKATOSA-N 0.000 description 1
- AIQWYVFNBNNOLU-RHYQMDGZSA-N Leu-Thr-Val Chemical compound [H]N[C@@H](CC(C)C)C(=O)N[C@@H]([C@@H](C)O)C(=O)N[C@@H](C(C)C)C(O)=O AIQWYVFNBNNOLU-RHYQMDGZSA-N 0.000 description 1
- UDXSLGLHFUBRRM-OEAJRASXSA-N Lys-Phe-Thr Chemical compound C[C@H]([C@@H](C(=O)O)NC(=O)[C@H](CC1=CC=CC=C1)NC(=O)[C@H](CCCCN)N)O UDXSLGLHFUBRRM-OEAJRASXSA-N 0.000 description 1
- LECIJRIRMVOFMH-ULQDDVLXSA-N Lys-Pro-Phe Chemical compound NCCCC[C@H](N)C(=O)N1CCC[C@H]1C(=O)N[C@H](C(O)=O)CC1=CC=CC=C1 LECIJRIRMVOFMH-ULQDDVLXSA-N 0.000 description 1
- KTINOHQFVVCEGQ-XIRDDKMYSA-N Lys-Trp-Asp Chemical compound NCCCC[C@H](N)C(=O)N[C@@H](Cc1c[nH]c2ccccc12)C(=O)N[C@@H](CC(O)=O)C(O)=O KTINOHQFVVCEGQ-XIRDDKMYSA-N 0.000 description 1
- NROQVSYLPRLJIP-PMVMPFDFSA-N Lys-Trp-Tyr Chemical compound [H]N[C@@H](CCCCN)C(=O)N[C@@H](CC1=CNC2=C1C=CC=C2)C(=O)N[C@@H](CC1=CC=C(O)C=C1)C(O)=O NROQVSYLPRLJIP-PMVMPFDFSA-N 0.000 description 1
- RMKJOQSYLQQRFN-KKUMJFAQSA-N Lys-Tyr-Asp Chemical compound [H]N[C@@H](CCCCN)C(=O)N[C@@H](CC1=CC=C(O)C=C1)C(=O)N[C@@H](CC(O)=O)C(O)=O RMKJOQSYLQQRFN-KKUMJFAQSA-N 0.000 description 1
- DRRXXZBXDMLGFC-IHRRRGAJSA-N Lys-Val-Leu Chemical compound CC(C)C[C@@H](C(O)=O)NC(=O)[C@H](C(C)C)NC(=O)[C@@H](N)CCCCN DRRXXZBXDMLGFC-IHRRRGAJSA-N 0.000 description 1
- SQUTUWHAAWJYES-GUBZILKMSA-N Met-Asp-Arg Chemical compound [H]N[C@@H](CCSC)C(=O)N[C@@H](CC(O)=O)C(=O)N[C@@H](CCCNC(N)=N)C(O)=O SQUTUWHAAWJYES-GUBZILKMSA-N 0.000 description 1
- FVKRQMQQFGBXHV-QXEWZRGKSA-N Met-Asp-Val Chemical compound CSCC[C@H](N)C(=O)N[C@@H](CC(O)=O)C(=O)N[C@@H](C(C)C)C(O)=O FVKRQMQQFGBXHV-QXEWZRGKSA-N 0.000 description 1
- JPCHYAUKOUGOIB-HJGDQZAQSA-N Met-Glu-Thr Chemical compound CSCC[C@H](N)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H]([C@@H](C)O)C(O)=O JPCHYAUKOUGOIB-HJGDQZAQSA-N 0.000 description 1
- VWWGEKCAPBMIFE-SRVKXCTJSA-N Met-Met-Met Chemical compound CSCC[C@H](N)C(=O)N[C@@H](CCSC)C(=O)N[C@@H](CCSC)C(O)=O VWWGEKCAPBMIFE-SRVKXCTJSA-N 0.000 description 1
- PNHRPOWKRRJATF-IHRRRGAJSA-N Met-Tyr-Ser Chemical compound CSCC[C@H](N)C(=O)N[C@H](C(=O)N[C@@H](CO)C(O)=O)CC1=CC=C(O)C=C1 PNHRPOWKRRJATF-IHRRRGAJSA-N 0.000 description 1
- IIHMNTBFPMRJCN-RCWTZXSCSA-N Met-Val-Thr Chemical compound CSCC[C@H](N)C(=O)N[C@@H](C(C)C)C(=O)N[C@@H]([C@@H](C)O)C(O)=O IIHMNTBFPMRJCN-RCWTZXSCSA-N 0.000 description 1
- 102000016943 Muramidase Human genes 0.000 description 1
- 108010014251 Muramidase Proteins 0.000 description 1
- 241001529936 Murinae Species 0.000 description 1
- OVRNDRQMDRJTHS-CBQIKETKSA-N N-Acetyl-D-Galactosamine Chemical compound CC(=O)N[C@H]1[C@@H](O)O[C@H](CO)[C@H](O)[C@@H]1O OVRNDRQMDRJTHS-CBQIKETKSA-N 0.000 description 1
- 108010062010 N-Acetylmuramoyl-L-alanine Amidase Proteins 0.000 description 1
- 125000003047 N-acetyl group Chemical group 0.000 description 1
- OVRNDRQMDRJTHS-RTRLPJTCSA-N N-acetyl-D-glucosamine Chemical compound CC(=O)N[C@H]1C(O)O[C@H](CO)[C@@H](O)[C@@H]1O OVRNDRQMDRJTHS-RTRLPJTCSA-N 0.000 description 1
- FDJKUWYYUZCUJX-AJKRCSPLSA-N N-glycoloyl-beta-neuraminic acid Chemical compound OC[C@@H](O)[C@@H](O)[C@@H]1O[C@](O)(C(O)=O)C[C@H](O)[C@H]1NC(=O)CO FDJKUWYYUZCUJX-AJKRCSPLSA-N 0.000 description 1
- SUHQNCLNRUAGOO-UHFFFAOYSA-N N-glycoloyl-neuraminic acid Natural products OCC(O)C(O)C(O)C(NC(=O)CO)C(O)CC(=O)C(O)=O SUHQNCLNRUAGOO-UHFFFAOYSA-N 0.000 description 1
- 108010002311 N-glycylglutamic acid Proteins 0.000 description 1
- CHJJGSNFBQVOTG-UHFFFAOYSA-N N-methyl-guanidine Natural products CNC(N)=N CHJJGSNFBQVOTG-UHFFFAOYSA-N 0.000 description 1
- 125000001429 N-terminal alpha-amino-acid group Chemical group 0.000 description 1
- BQVUABVGYYSDCJ-UHFFFAOYSA-N Nalpha-L-Leucyl-L-tryptophan Natural products C1=CC=C2C(CC(NC(=O)C(N)CC(C)C)C(O)=O)=CNC2=C1 BQVUABVGYYSDCJ-UHFFFAOYSA-N 0.000 description 1
- 101100342977 Neurospora crassa (strain ATCC 24698 / 74-OR23-1A / CBS 708.71 / DSM 1257 / FGSC 987) leu-1 gene Proteins 0.000 description 1
- 230000004989 O-glycosylation Effects 0.000 description 1
- BZQFBWGGLXLEPQ-UHFFFAOYSA-N O-phosphoryl-L-serine Natural products OC(=O)C(N)COP(O)(O)=O BZQFBWGGLXLEPQ-UHFFFAOYSA-N 0.000 description 1
- 108091034117 Oligonucleotide Proteins 0.000 description 1
- 238000012408 PCR amplification Methods 0.000 description 1
- 229910019142 PO4 Inorganic materials 0.000 description 1
- YYRCPTVAPLQRNC-ULQDDVLXSA-N Phe-Arg-Lys Chemical compound NCCCC[C@@H](C(O)=O)NC(=O)[C@H](CCCNC(N)=N)NC(=O)[C@@H](N)CC1=CC=CC=C1 YYRCPTVAPLQRNC-ULQDDVLXSA-N 0.000 description 1
- LJUUGSWZPQOJKD-JYJNAYRXSA-N Phe-Arg-Val Chemical compound CC(C)[C@H](NC(=O)[C@H](CCCNC(N)=N)NC(=O)[C@@H](N)Cc1ccccc1)C(O)=O LJUUGSWZPQOJKD-JYJNAYRXSA-N 0.000 description 1
- BXNGIHFNNNSEOS-UWVGGRQHSA-N Phe-Asn Chemical compound NC(=O)C[C@@H](C(O)=O)NC(=O)[C@@H](N)CC1=CC=CC=C1 BXNGIHFNNNSEOS-UWVGGRQHSA-N 0.000 description 1
- JJHVFCUWLSKADD-ONGXEEELSA-N Phe-Gly-Ala Chemical compound [H]N[C@@H](CC1=CC=CC=C1)C(=O)NCC(=O)N[C@@H](C)C(O)=O JJHVFCUWLSKADD-ONGXEEELSA-N 0.000 description 1
- MMJJFXWMCMJMQA-STQMWFEESA-N Phe-Pro-Gly Chemical compound C([C@H](N)C(=O)N1[C@@H](CCC1)C(=O)NCC(O)=O)C1=CC=CC=C1 MMJJFXWMCMJMQA-STQMWFEESA-N 0.000 description 1
- GNZCMRRSXOBHLC-JYJNAYRXSA-N Phe-Val-Met Chemical compound CC(C)[C@@H](C(=O)N[C@@H](CCSC)C(=O)O)NC(=O)[C@H](CC1=CC=CC=C1)N GNZCMRRSXOBHLC-JYJNAYRXSA-N 0.000 description 1
- 101710182846 Polyhedrin Proteins 0.000 description 1
- 108010039918 Polylysine Proteins 0.000 description 1
- 101800001357 Potential peptide Proteins 0.000 description 1
- 102400000745 Potential peptide Human genes 0.000 description 1
- GLEOIKLQBZNKJZ-WDSKDSINSA-N Pro-Asp Chemical compound OC(=O)C[C@@H](C(O)=O)NC(=O)[C@@H]1CCCN1 GLEOIKLQBZNKJZ-WDSKDSINSA-N 0.000 description 1
- HXOLCSYHGRNXJJ-IHRRRGAJSA-N Pro-Asp-Phe Chemical compound [H]N1CCC[C@H]1C(=O)N[C@@H](CC(O)=O)C(=O)N[C@@H](CC1=CC=CC=C1)C(O)=O HXOLCSYHGRNXJJ-IHRRRGAJSA-N 0.000 description 1
- MGDFPGCFVJFITQ-CIUDSAMLSA-N Pro-Glu-Asp Chemical compound [H]N1CCC[C@H]1C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CC(O)=O)C(O)=O MGDFPGCFVJFITQ-CIUDSAMLSA-N 0.000 description 1
- NXEYSLRNNPWCRN-SRVKXCTJSA-N Pro-Glu-Leu Chemical compound [H]N1CCC[C@H]1C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CC(C)C)C(O)=O NXEYSLRNNPWCRN-SRVKXCTJSA-N 0.000 description 1
- HAAQQNHQZBOWFO-LURJTMIESA-N Pro-Gly-Gly Chemical compound OC(=O)CNC(=O)CNC(=O)[C@@H]1CCCN1 HAAQQNHQZBOWFO-LURJTMIESA-N 0.000 description 1
- DXTOOBDIIAJZBJ-BQBZGAKWSA-N Pro-Gly-Ser Chemical compound [H]N1CCC[C@H]1C(=O)NCC(=O)N[C@@H](CO)C(O)=O DXTOOBDIIAJZBJ-BQBZGAKWSA-N 0.000 description 1
- XYSXOCIWCPFOCG-IHRRRGAJSA-N Pro-Leu-Leu Chemical compound [H]N1CCC[C@H]1C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](CC(C)C)C(O)=O XYSXOCIWCPFOCG-IHRRRGAJSA-N 0.000 description 1
- SWRNSCMUXRLHCR-ULQDDVLXSA-N Pro-Phe-Lys Chemical compound C([C@@H](C(=O)N[C@@H](CCCCN)C(O)=O)NC(=O)[C@H]1NCCC1)C1=CC=CC=C1 SWRNSCMUXRLHCR-ULQDDVLXSA-N 0.000 description 1
- POQFNPILEQEODH-FXQIFTODSA-N Pro-Ser-Ala Chemical compound [H]N1CCC[C@H]1C(=O)N[C@@H](CO)C(=O)N[C@@H](C)C(O)=O POQFNPILEQEODH-FXQIFTODSA-N 0.000 description 1
- AIOWVDNPESPXRB-YTWAJWBKSA-N Pro-Thr-Pro Chemical compound C[C@H]([C@@H](C(=O)N1CCC[C@@H]1C(=O)O)NC(=O)[C@@H]2CCCN2)O AIOWVDNPESPXRB-YTWAJWBKSA-N 0.000 description 1
- XDKKMRPRRCOELJ-GUBZILKMSA-N Pro-Val-Ala Chemical compound OC(=O)[C@H](C)NC(=O)[C@H](C(C)C)NC(=O)[C@@H]1CCCN1 XDKKMRPRRCOELJ-GUBZILKMSA-N 0.000 description 1
- ONIBWKKTOPOVIA-UHFFFAOYSA-N Proline Natural products OC(=O)C1CCCN1 ONIBWKKTOPOVIA-UHFFFAOYSA-N 0.000 description 1
- 239000012722 SDS sample buffer Substances 0.000 description 1
- HRNQLKCLPVKZNE-CIUDSAMLSA-N Ser-Ala-Leu Chemical compound [H]N[C@@H](CO)C(=O)N[C@@H](C)C(=O)N[C@@H](CC(C)C)C(O)=O HRNQLKCLPVKZNE-CIUDSAMLSA-N 0.000 description 1
- UBRXAVQWXOWRSJ-ZLUOBGJFSA-N Ser-Asn-Asp Chemical compound C([C@@H](C(=O)N[C@@H](CC(=O)O)C(=O)O)NC(=O)[C@H](CO)N)C(=O)N UBRXAVQWXOWRSJ-ZLUOBGJFSA-N 0.000 description 1
- FIDMVVBUOCMMJG-CIUDSAMLSA-N Ser-Asn-Leu Chemical compound CC(C)C[C@@H](C(O)=O)NC(=O)[C@H](CC(N)=O)NC(=O)[C@@H](N)CO FIDMVVBUOCMMJG-CIUDSAMLSA-N 0.000 description 1
- YMTLKLXDFCSCNX-BYPYZUCNSA-N Ser-Gly-Gly Chemical compound OC[C@H](N)C(=O)NCC(=O)NCC(O)=O YMTLKLXDFCSCNX-BYPYZUCNSA-N 0.000 description 1
- SFTZWNJFZYOLBD-ZDLURKLDSA-N Ser-Gly-Thr Chemical compound C[C@@H](O)[C@@H](C(O)=O)NC(=O)CNC(=O)[C@@H](N)CO SFTZWNJFZYOLBD-ZDLURKLDSA-N 0.000 description 1
- LOKXAXAESFYFAX-CIUDSAMLSA-N Ser-His-Cys Chemical compound OC[C@H](N)C(=O)N[C@H](C(=O)N[C@@H](CS)C(O)=O)CC1=CN=CN1 LOKXAXAESFYFAX-CIUDSAMLSA-N 0.000 description 1
- NFDYGNFETJVMSE-BQBZGAKWSA-N Ser-Leu Chemical compound CC(C)C[C@@H](C(O)=O)NC(=O)[C@@H](N)CO NFDYGNFETJVMSE-BQBZGAKWSA-N 0.000 description 1
- WBAXJMCUFIXCNI-WDSKDSINSA-N Ser-Pro Chemical compound OC[C@H](N)C(=O)N1CCC[C@H]1C(O)=O WBAXJMCUFIXCNI-WDSKDSINSA-N 0.000 description 1
- DINQYZRMXGWWTG-GUBZILKMSA-N Ser-Pro-Pro Chemical compound OC[C@H](N)C(=O)N1CCC[C@H]1C(=O)N1[C@H](C(O)=O)CCC1 DINQYZRMXGWWTG-GUBZILKMSA-N 0.000 description 1
- LGIMRDKGABDMBN-DCAQKATOSA-N Ser-Val-Lys Chemical compound CC(C)[C@@H](C(=O)N[C@@H](CCCCN)C(=O)O)NC(=O)[C@H](CO)N LGIMRDKGABDMBN-DCAQKATOSA-N 0.000 description 1
- BQCADISMDOOEFD-UHFFFAOYSA-N Silver Chemical compound [Ag] BQCADISMDOOEFD-UHFFFAOYSA-N 0.000 description 1
- PKXHGEXFMIZSER-QTKMDUPCSA-N Thr-Arg-His Chemical compound C[C@H]([C@@H](C(=O)N[C@@H](CCCN=C(N)N)C(=O)N[C@@H](CC1=CN=CN1)C(=O)O)N)O PKXHGEXFMIZSER-QTKMDUPCSA-N 0.000 description 1
- IQHUITKNHOKGFC-MIMYLULJSA-N Thr-Phe Chemical compound C[C@@H](O)[C@H](N)C(=O)N[C@H](C(O)=O)CC1=CC=CC=C1 IQHUITKNHOKGFC-MIMYLULJSA-N 0.000 description 1
- WNQJTLATMXYSEL-OEAJRASXSA-N Thr-Phe-Leu Chemical compound [H]N[C@@H]([C@@H](C)O)C(=O)N[C@@H](CC1=CC=CC=C1)C(=O)N[C@@H](CC(C)C)C(O)=O WNQJTLATMXYSEL-OEAJRASXSA-N 0.000 description 1
- BDENGIGFTNYZSJ-RCWTZXSCSA-N Thr-Pro-Met Chemical compound [H]N[C@@H]([C@@H](C)O)C(=O)N1CCC[C@H]1C(=O)N[C@@H](CCSC)C(O)=O BDENGIGFTNYZSJ-RCWTZXSCSA-N 0.000 description 1
- GVMXJJAJLIEASL-ZJDVBMNYSA-N Thr-Pro-Thr Chemical compound C[C@@H](O)[C@H](N)C(=O)N1CCC[C@H]1C(=O)N[C@@H]([C@@H](C)O)C(O)=O GVMXJJAJLIEASL-ZJDVBMNYSA-N 0.000 description 1
- PWIQCLSQVQBOQV-AAEUAGOBSA-N Trp-Glu Chemical compound C1=CC=C2C(C[C@H](N)C(=O)N[C@@H](CCC(O)=O)C(O)=O)=CNC2=C1 PWIQCLSQVQBOQV-AAEUAGOBSA-N 0.000 description 1
- JVTHMUDOKPQBOT-NSHDSACASA-N Trp-Gly-Gly Chemical compound C1=CC=C2C(C[C@H]([NH3+])C(=O)NCC(=O)NCC([O-])=O)=CNC2=C1 JVTHMUDOKPQBOT-NSHDSACASA-N 0.000 description 1
- VPRHDRKAPYZMHL-SZMVWBNQSA-N Trp-Leu-Glu Chemical compound C1=CC=C2C(C[C@H](N)C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](CCC(O)=O)C(O)=O)=CNC2=C1 VPRHDRKAPYZMHL-SZMVWBNQSA-N 0.000 description 1
- DVLHKUWLNKDINO-PMVMPFDFSA-N Trp-Tyr-Leu Chemical compound [H]N[C@@H](CC1=CNC2=C1C=CC=C2)C(=O)N[C@@H](CC1=CC=C(O)C=C1)C(=O)N[C@@H](CC(C)C)C(O)=O DVLHKUWLNKDINO-PMVMPFDFSA-N 0.000 description 1
- QIVBCDIJIAJPQS-UHFFFAOYSA-N Tryptophan Natural products C1=CC=C2C(CC(N)C(O)=O)=CNC2=C1 QIVBCDIJIAJPQS-UHFFFAOYSA-N 0.000 description 1
- 206010054094 Tumour necrosis Diseases 0.000 description 1
- CDHQEOXPWBDFPL-QWRGUYRKSA-N Tyr-Gly-Asn Chemical compound NC(=O)C[C@@H](C(O)=O)NC(=O)CNC(=O)[C@@H](N)CC1=CC=C(O)C=C1 CDHQEOXPWBDFPL-QWRGUYRKSA-N 0.000 description 1
- AVFGBGGRZOKSFS-KJEVXHAQSA-N Tyr-Met-Thr Chemical compound C[C@H]([C@@H](C(=O)O)NC(=O)[C@H](CCSC)NC(=O)[C@H](CC1=CC=C(C=C1)O)N)O AVFGBGGRZOKSFS-KJEVXHAQSA-N 0.000 description 1
- SOAUMCDLIUGXJJ-SRVKXCTJSA-N Tyr-Ser-Asn Chemical compound [H]N[C@@H](CC1=CC=C(O)C=C1)C(=O)N[C@@H](CO)C(=O)N[C@@H](CC(N)=O)C(O)=O SOAUMCDLIUGXJJ-SRVKXCTJSA-N 0.000 description 1
- LDKDSFQSEUOCOO-RPTUDFQQSA-N Tyr-Thr-Phe Chemical compound [H]N[C@@H](CC1=CC=C(O)C=C1)C(=O)N[C@@H]([C@@H](C)O)C(=O)N[C@@H](CC1=CC=CC=C1)C(O)=O LDKDSFQSEUOCOO-RPTUDFQQSA-N 0.000 description 1
- HSCJRCZFDFQWRP-ABVWGUQPSA-N UDP-alpha-D-galactose Chemical compound O[C@@H]1[C@@H](O)[C@@H](O)[C@@H](CO)O[C@@H]1OP(O)(=O)OP(O)(=O)OC[C@@H]1[C@@H](O)[C@@H](O)[C@H](N2C(NC(=O)C=C2)=O)O1 HSCJRCZFDFQWRP-ABVWGUQPSA-N 0.000 description 1
- HSCJRCZFDFQWRP-UHFFFAOYSA-N Uridindiphosphoglukose Natural products OC1C(O)C(O)C(CO)OC1OP(O)(=O)OP(O)(=O)OCC1C(O)C(O)C(N2C(NC(=O)C=C2)=O)O1 HSCJRCZFDFQWRP-UHFFFAOYSA-N 0.000 description 1
- YFOCMOVJBQDBCE-NRPADANISA-N Val-Ala-Glu Chemical compound C[C@@H](C(=O)N[C@@H](CCC(=O)O)C(=O)O)NC(=O)[C@H](C(C)C)N YFOCMOVJBQDBCE-NRPADANISA-N 0.000 description 1
- UEHRGZCNLSWGHK-DLOVCJGASA-N Val-Glu-Val Chemical compound CC(C)[C@H](N)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](C(C)C)C(O)=O UEHRGZCNLSWGHK-DLOVCJGASA-N 0.000 description 1
- OTJMMKPMLUNTQT-AVGNSLFASA-N Val-Leu-Arg Chemical compound CC(C)C[C@@H](C(=O)N[C@@H](CCCN=C(N)N)C(=O)O)NC(=O)[C@H](C(C)C)N OTJMMKPMLUNTQT-AVGNSLFASA-N 0.000 description 1
- LJSZPMSUYKKKCP-UBHSHLNASA-N Val-Phe-Ala Chemical compound CC(C)[C@H](N)C(=O)N[C@H](C(=O)N[C@@H](C)C(O)=O)CC1=CC=CC=C1 LJSZPMSUYKKKCP-UBHSHLNASA-N 0.000 description 1
- ZXYPHBKIZLAQTL-QXEWZRGKSA-N Val-Pro-Asp Chemical compound CC(C)[C@@H](C(=O)N1CCC[C@H]1C(=O)N[C@@H](CC(=O)O)C(=O)O)N ZXYPHBKIZLAQTL-QXEWZRGKSA-N 0.000 description 1
- QWCZXKIFPWPQHR-JYJNAYRXSA-N Val-Pro-Tyr Chemical compound CC(C)[C@H](N)C(=O)N1CCC[C@H]1C(=O)N[C@H](C(O)=O)CC1=CC=C(O)C=C1 QWCZXKIFPWPQHR-JYJNAYRXSA-N 0.000 description 1
- GBIUHAYJGWVNLN-AEJSXWLSSA-N Val-Ser-Pro Chemical compound CC(C)[C@@H](C(=O)N[C@@H](CO)C(=O)N1CCC[C@@H]1C(=O)O)N GBIUHAYJGWVNLN-AEJSXWLSSA-N 0.000 description 1
- NZYNRRGJJVSSTJ-GUBZILKMSA-N Val-Ser-Val Chemical compound CC(C)[C@H](N)C(=O)N[C@@H](CO)C(=O)N[C@@H](C(C)C)C(O)=O NZYNRRGJJVSSTJ-GUBZILKMSA-N 0.000 description 1
- PFMSJVIPEZMKSC-DZKIICNBSA-N Val-Tyr-Glu Chemical compound CC(C)[C@@H](C(=O)N[C@@H](CC1=CC=C(C=C1)O)C(=O)N[C@@H](CCC(=O)O)C(=O)O)N PFMSJVIPEZMKSC-DZKIICNBSA-N 0.000 description 1
- AEFJNECXZCODJM-UWVGGRQHSA-N Val-Val-Gly Chemical compound CC(C)[C@H]([NH3+])C(=O)N[C@@H](C(C)C)C(=O)NCC([O-])=O AEFJNECXZCODJM-UWVGGRQHSA-N 0.000 description 1
- 241000464917 Vieja Species 0.000 description 1
- 238000010521 absorption reaction Methods 0.000 description 1
- 125000000218 acetic acid group Chemical group C(C)(=O)* 0.000 description 1
- 230000021736 acetylation Effects 0.000 description 1
- 238000006640 acetylation reaction Methods 0.000 description 1
- 229960000643 adenine Drugs 0.000 description 1
- 229960005305 adenosine Drugs 0.000 description 1
- 108010087924 alanylproline Proteins 0.000 description 1
- ZTOKCBJDEGPICW-GWPISINRSA-N alpha-D-Manp-(1->3)-[alpha-D-Manp-(1->6)]-beta-D-Manp-(1->4)-beta-D-GlcpNAc-(1->4)-beta-D-GlcpNAc Chemical compound O[C@@H]1[C@@H](NC(=O)C)[C@H](O)O[C@H](CO)[C@H]1O[C@H]1[C@H](NC(C)=O)[C@@H](O)[C@H](O[C@H]2[C@H]([C@@H](O[C@@H]3[C@H]([C@@H](O)[C@H](O)[C@@H](CO)O3)O)[C@H](O)[C@@H](CO[C@@H]3[C@H]([C@@H](O)[C@H](O)[C@@H](CO)O3)O)O2)O)[C@@H](CO)O1 ZTOKCBJDEGPICW-GWPISINRSA-N 0.000 description 1
- WQZGKKKJIJFFOK-PHYPRBDBSA-N alpha-D-galactose Chemical compound OC[C@H]1O[C@H](O)[C@H](O)[C@@H](O)[C@H]1O WQZGKKKJIJFFOK-PHYPRBDBSA-N 0.000 description 1
- 230000003321 amplification Effects 0.000 description 1
- 239000003957 anion exchange resin Substances 0.000 description 1
- 238000000137 annealing Methods 0.000 description 1
- 230000000259 anti-tumor effect Effects 0.000 description 1
- 229940088710 antibiotic agent Drugs 0.000 description 1
- PYMYPHUHKUWMLA-UHFFFAOYSA-N arabinose Natural products OCC(O)C(O)C(O)C=O PYMYPHUHKUWMLA-UHFFFAOYSA-N 0.000 description 1
- 108010068265 aspartyltyrosine Proteins 0.000 description 1
- 230000010310 bacterial transformation Effects 0.000 description 1
- 230000008901 benefit Effects 0.000 description 1
- SRBFZHDQGSBBOR-UHFFFAOYSA-N beta-D-Pyranose-Lyxose Natural products OC1COC(O)C(O)C1O SRBFZHDQGSBBOR-UHFFFAOYSA-N 0.000 description 1
- AEMOLEFTQBMNLQ-UHFFFAOYSA-N beta-D-galactopyranuronic acid Natural products OC1OC(C(O)=O)C(O)C(O)C1O AEMOLEFTQBMNLQ-UHFFFAOYSA-N 0.000 description 1
- IQFYYKKMVGJFEH-UHFFFAOYSA-N beta-L-thymidine Natural products O=C1NC(=O)C(C)=CN1C1OC(CO)C(O)C1 IQFYYKKMVGJFEH-UHFFFAOYSA-N 0.000 description 1
- DRTQHJPVMGBUCF-PSQAKQOGSA-N beta-L-uridine Natural products O[C@H]1[C@@H](O)[C@H](CO)O[C@@H]1N1C(=O)NC(=O)C=C1 DRTQHJPVMGBUCF-PSQAKQOGSA-N 0.000 description 1
- 230000008033 biological extinction Effects 0.000 description 1
- 230000031018 biological processes and functions Effects 0.000 description 1
- 229960000182 blood factors Drugs 0.000 description 1
- 229910052799 carbon Inorganic materials 0.000 description 1
- 230000021523 carboxylation Effects 0.000 description 1
- 238000006473 carboxylation reaction Methods 0.000 description 1
- 238000006555 catalytic reaction Methods 0.000 description 1
- 238000004113 cell culture Methods 0.000 description 1
- 230000003196 chaotropic effect Effects 0.000 description 1
- 238000012512 characterization method Methods 0.000 description 1
- 150000005829 chemical entities Chemical class 0.000 description 1
- 230000000295 complement effect Effects 0.000 description 1
- 238000003271 compound fluorescence assay Methods 0.000 description 1
- 239000000470 constituent Substances 0.000 description 1
- 238000010276 construction Methods 0.000 description 1
- 239000000287 crude extract Substances 0.000 description 1
- 238000012258 culturing Methods 0.000 description 1
- OOTFVKOQINZBBF-UHFFFAOYSA-N cystamine Chemical compound CCSSCCN OOTFVKOQINZBBF-UHFFFAOYSA-N 0.000 description 1
- 229940099500 cystamine Drugs 0.000 description 1
- UFULAYFCSOUIOV-UHFFFAOYSA-N cysteamine Chemical compound NCCS UFULAYFCSOUIOV-UHFFFAOYSA-N 0.000 description 1
- 108010016616 cysteinylglycine Proteins 0.000 description 1
- 229960003067 cystine Drugs 0.000 description 1
- UHDGCWIWMRVCDJ-ZAKLUEHWSA-N cytidine Chemical compound O=C1N=C(N)C=CN1[C@H]1[C@H](O)[C@@H](O)[C@H](CO)O1 UHDGCWIWMRVCDJ-ZAKLUEHWSA-N 0.000 description 1
- IERHLVCPSMICTF-XVFCMESISA-N cytidine 5'-monophosphate Chemical compound O=C1N=C(N)C=CN1[C@H]1[C@H](O)[C@H](O)[C@@H](COP(O)(O)=O)O1 IERHLVCPSMICTF-XVFCMESISA-N 0.000 description 1
- IERHLVCPSMICTF-UHFFFAOYSA-N cytidine monophosphate Natural products O=C1N=C(N)C=CN1C1C(O)C(O)C(COP(O)(O)=O)O1 IERHLVCPSMICTF-UHFFFAOYSA-N 0.000 description 1
- 230000009089 cytolysis Effects 0.000 description 1
- 229940104302 cytosine Drugs 0.000 description 1
- 238000004925 denaturation Methods 0.000 description 1
- 230000036425 denaturation Effects 0.000 description 1
- 238000001212 derivatisation Methods 0.000 description 1
- 238000013461 design Methods 0.000 description 1
- 230000000368 destabilizing effect Effects 0.000 description 1
- 238000001514 detection method Methods 0.000 description 1
- 239000003599 detergent Substances 0.000 description 1
- 229950006137 dexfosfoserine Drugs 0.000 description 1
- 239000012954 diazonium Substances 0.000 description 1
- 150000001989 diazonium salts Chemical class 0.000 description 1
- SWSQBOPZIKWTGO-UHFFFAOYSA-N dimethylaminoamidine Natural products CN(C)C(N)=N SWSQBOPZIKWTGO-UHFFFAOYSA-N 0.000 description 1
- 239000012153 distilled water Substances 0.000 description 1
- PMMYEEVYMWASQN-UHFFFAOYSA-N dl-hydroxyproline Natural products OC1C[NH2+]C(C([O-])=O)C1 PMMYEEVYMWASQN-UHFFFAOYSA-N 0.000 description 1
- 239000003623 enhancer Substances 0.000 description 1
- 230000002708 enhancing effect Effects 0.000 description 1
- 238000006911 enzymatic reaction Methods 0.000 description 1
- 238000001952 enzyme assay Methods 0.000 description 1
- 238000001976 enzyme digestion Methods 0.000 description 1
- 238000000605 extraction Methods 0.000 description 1
- 230000002349 favourable effect Effects 0.000 description 1
- 210000002950 fibroblast Anatomy 0.000 description 1
- 235000013305 food Nutrition 0.000 description 1
- 229930182830 galactose Natural products 0.000 description 1
- 238000001502 gel electrophoresis Methods 0.000 description 1
- 238000010353 genetic engineering Methods 0.000 description 1
- 239000003862 glucocorticoid Substances 0.000 description 1
- 229960002442 glucosamine Drugs 0.000 description 1
- 239000008103 glucose Substances 0.000 description 1
- 229940097043 glucuronic acid Drugs 0.000 description 1
- 235000013922 glutamic acid Nutrition 0.000 description 1
- 239000004220 glutamic acid Substances 0.000 description 1
- 108010049041 glutamylalanine Proteins 0.000 description 1
- 229960003180 glutathione Drugs 0.000 description 1
- 230000001279 glycosylating effect Effects 0.000 description 1
- 108010037850 glycylvaline Proteins 0.000 description 1
- 210000002288 golgi apparatus Anatomy 0.000 description 1
- 239000011544 gradient gel Substances 0.000 description 1
- 230000005484 gravity Effects 0.000 description 1
- 230000012010 growth Effects 0.000 description 1
- 229960000789 guanidine hydrochloride Drugs 0.000 description 1
- 229940029575 guanosine Drugs 0.000 description 1
- 238000004128 high performance liquid chromatography Methods 0.000 description 1
- 108010085325 histidylproline Proteins 0.000 description 1
- 230000001744 histochemical effect Effects 0.000 description 1
- 125000004356 hydroxy functional group Chemical group O* 0.000 description 1
- 229960002591 hydroxyproline Drugs 0.000 description 1
- 230000001900 immune effect Effects 0.000 description 1
- 230000002163 immunogen Effects 0.000 description 1
- 230000005847 immunogenicity Effects 0.000 description 1
- 230000001976 improved effect Effects 0.000 description 1
- 238000010348 incorporation Methods 0.000 description 1
- 230000001939 inductive effect Effects 0.000 description 1
- 230000000977 initiatory effect Effects 0.000 description 1
- 230000010354 integration Effects 0.000 description 1
- 229940079322 interferon Drugs 0.000 description 1
- 229960003130 interferon gamma Drugs 0.000 description 1
- 229960001388 interferon-beta Drugs 0.000 description 1
- 230000003834 intracellular effect Effects 0.000 description 1
- 238000004255 ion exchange chromatography Methods 0.000 description 1
- AGPKZVBTJJNPAG-UHFFFAOYSA-N isoleucine Natural products CCC(C)C(N)C(O)=O AGPKZVBTJJNPAG-UHFFFAOYSA-N 0.000 description 1
- 229960000310 isoleucine Drugs 0.000 description 1
- 210000003734 kidney Anatomy 0.000 description 1
- 238000011031 large-scale manufacturing process Methods 0.000 description 1
- 238000001638 lipofection Methods 0.000 description 1
- 229960000274 lysozyme Drugs 0.000 description 1
- 235000010335 lysozyme Nutrition 0.000 description 1
- 239000004325 lysozyme Substances 0.000 description 1
- 108010017391 lysylvaline Proteins 0.000 description 1
- 230000014759 maintenance of location Effects 0.000 description 1
- 238000004949 mass spectrometry Methods 0.000 description 1
- 239000012528 membrane Substances 0.000 description 1
- 229960003151 mercaptamine Drugs 0.000 description 1
- 230000004060 metabolic process Effects 0.000 description 1
- 229910052751 metal Inorganic materials 0.000 description 1
- 239000002184 metal Substances 0.000 description 1
- 150000002739 metals Chemical class 0.000 description 1
- MYWUZJCMWCOHBA-VIFPVBQESA-N methamphetamine Chemical compound CN[C@@H](C)CC1=CC=CC=C1 MYWUZJCMWCOHBA-VIFPVBQESA-N 0.000 description 1
- 229930182817 methionine Natural products 0.000 description 1
- 229960000485 methotrexate Drugs 0.000 description 1
- 238000012544 monitoring process Methods 0.000 description 1
- 229930027945 nicotinamide-adenine dinucleotide Natural products 0.000 description 1
- BOPGDPNILDQYTO-NNYOXOHSSA-N nicotinamide-adenine dinucleotide Chemical compound C1=CCC(C(=O)N)=CN1[C@H]1[C@H](O)[C@H](O)[C@@H](COP(O)(=O)OP(O)(=O)OC[C@@H]2[C@H]([C@@H](O)[C@@H](O2)N2C3=NC=NC(N)=C3N=C2)O)O1 BOPGDPNILDQYTO-NNYOXOHSSA-N 0.000 description 1
- 238000003199 nucleic acid amplification method Methods 0.000 description 1
- 238000005457 optimization Methods 0.000 description 1
- 210000001672 ovary Anatomy 0.000 description 1
- YPZRWBKMTBYPTK-UHFFFAOYSA-N oxidized gamma-L-glutamyl-L-cysteinylglycine Natural products OC(=O)C(N)CCC(=O)NC(C(=O)NCC(O)=O)CSSCC(C(=O)NCC(O)=O)NC(=O)CCC(N)C(O)=O YPZRWBKMTBYPTK-UHFFFAOYSA-N 0.000 description 1
- 239000000813 peptide hormone Substances 0.000 description 1
- 230000002093 peripheral effect Effects 0.000 description 1
- 239000012466 permeate Substances 0.000 description 1
- NBIIXXVUZAFLBC-UHFFFAOYSA-K phosphate Chemical compound [O-]P([O-])([O-])=O NBIIXXVUZAFLBC-UHFFFAOYSA-K 0.000 description 1
- 239000010452 phosphate Substances 0.000 description 1
- BZQFBWGGLXLEPQ-REOHCLBHSA-N phosphoserine Chemical compound OC(=O)[C@@H](N)COP(O)(O)=O BZQFBWGGLXLEPQ-REOHCLBHSA-N 0.000 description 1
- USRGIUJOYOXOQJ-GBXIJSLDSA-N phosphothreonine Chemical compound OP(=O)(O)O[C@H](C)[C@H](N)C(O)=O USRGIUJOYOXOQJ-GBXIJSLDSA-N 0.000 description 1
- DCWXELXMIBXGTH-UHFFFAOYSA-N phosphotyrosine Chemical group OC(=O)C(N)CC1=CC=C(OP(O)(O)=O)C=C1 DCWXELXMIBXGTH-UHFFFAOYSA-N 0.000 description 1
- 230000001766 physiological effect Effects 0.000 description 1
- 239000013600 plasmid vector Substances 0.000 description 1
- 229920001467 poly(styrenesulfonates) Polymers 0.000 description 1
- 238000002264 polyacrylamide gel electrophoresis Methods 0.000 description 1
- 229920000656 polylysine Polymers 0.000 description 1
- 229920001451 polypropylene glycol Polymers 0.000 description 1
- 125000002924 primary amino group Chemical group [H]N([H])* 0.000 description 1
- 108010029020 prolylglycine Proteins 0.000 description 1
- 230000030788 protein refolding Effects 0.000 description 1
- 238000001799 protein solubilization Methods 0.000 description 1
- 230000007925 protein solubilization Effects 0.000 description 1
- 230000012743 protein tagging Effects 0.000 description 1
- 230000006337 proteolytic cleavage Effects 0.000 description 1
- 238000011084 recovery Methods 0.000 description 1
- 230000009467 reduction Effects 0.000 description 1
- 239000004627 regenerated cellulose Substances 0.000 description 1
- 238000007634 remodeling Methods 0.000 description 1
- 230000010076 replication Effects 0.000 description 1
- 238000012827 research and development Methods 0.000 description 1
- 230000004044 response Effects 0.000 description 1
- 230000000284 resting effect Effects 0.000 description 1
- 239000011369 resultant mixture Substances 0.000 description 1
- 239000012465 retentate Substances 0.000 description 1
- 230000001177 retroviral effect Effects 0.000 description 1
- 238000004007 reversed phase HPLC Methods 0.000 description 1
- 238000012552 review Methods 0.000 description 1
- 238000007790 scraping Methods 0.000 description 1
- 238000000926 separation method Methods 0.000 description 1
- 230000009450 sialylation Effects 0.000 description 1
- 229910052709 silver Inorganic materials 0.000 description 1
- 239000004332 silver Substances 0.000 description 1
- 238000001542 size-exclusion chromatography Methods 0.000 description 1
- 150000003384 small molecules Chemical class 0.000 description 1
- 239000011537 solubilization buffer Substances 0.000 description 1
- 238000003153 stable transfection Methods 0.000 description 1
- 238000010186 staining Methods 0.000 description 1
- 239000003774 sulfhydryl reagent Substances 0.000 description 1
- 229940037128 systemic glucocorticoids Drugs 0.000 description 1
- 230000008685 targeting Effects 0.000 description 1
- 229940124597 therapeutic agent Drugs 0.000 description 1
- 229940104230 thymidine Drugs 0.000 description 1
- 229940113082 thymine Drugs 0.000 description 1
- FGMPLJWBKKVCDB-UHFFFAOYSA-N trans-L-hydroxy-proline Natural products ON1CCCC1C(O)=O FGMPLJWBKKVCDB-UHFFFAOYSA-N 0.000 description 1
- 230000002103 transcriptional effect Effects 0.000 description 1
- 238000006276 transfer reaction Methods 0.000 description 1
- 230000032258 transport Effects 0.000 description 1
- 229940035893 uracil Drugs 0.000 description 1
- DRTQHJPVMGBUCF-UHFFFAOYSA-N uracil arabinoside Natural products OC1C(O)C(CO)OC1N1C(=O)NC(=O)C=C1 DRTQHJPVMGBUCF-UHFFFAOYSA-N 0.000 description 1
- 229940045145 uridine Drugs 0.000 description 1
- 239000004474 valine Substances 0.000 description 1
- 108700026220 vif Genes Proteins 0.000 description 1
- 239000013603 viral vector Substances 0.000 description 1
- 239000012130 whole-cell lysate Substances 0.000 description 1
Classifications
-
- C—CHEMISTRY; METALLURGY
- C12—BIOCHEMISTRY; BEER; SPIRITS; WINE; VINEGAR; MICROBIOLOGY; ENZYMOLOGY; MUTATION OR GENETIC ENGINEERING
- C12N—MICROORGANISMS OR ENZYMES; COMPOSITIONS THEREOF; PROPAGATING, PRESERVING, OR MAINTAINING MICROORGANISMS; MUTATION OR GENETIC ENGINEERING; CULTURE MEDIA
- C12N9/00—Enzymes; Proenzymes; Compositions thereof; Processes for preparing, activating, inhibiting, separating or purifying enzymes
- C12N9/10—Transferases (2.)
- C12N9/1048—Glycosyltransferases (2.4)
- C12N9/1051—Hexosyltransferases (2.4.1)
-
- A—HUMAN NECESSITIES
- A61—MEDICAL OR VETERINARY SCIENCE; HYGIENE
- A61P—SPECIFIC THERAPEUTIC ACTIVITY OF CHEMICAL COMPOUNDS OR MEDICINAL PREPARATIONS
- A61P43/00—Drugs for specific purposes, not provided for in groups A61P1/00-A61P41/00
Definitions
- the present invention features compositions and methods related to truncated mutants of GalNAcT2.
- the invention features truncated human GalNAcT2 polypeptides.
- the invention also features nucleic acids encoding such truncated polypeptides, as well as vectors, host cells, expression systems, and methods of expressing and using such polypeptides.
- glycosyltransferases catalyze the synthesis of glycolipids, glycopeptides, and polysaccharides, by transferring an activated mono- or oligosaccharide residue to an existing acceptor molecule for the initiation or elongation of the carbohydrate chain.
- a catalytic reaction is believed to involve the recognition of both the donor and acceptor by suitable domains, as well as the catalytic site of the enzyme.
- peptide therapeutics are glycosylated peptides.
- the production of a recombinant glycopeptide as opposed to a recombinant non-glycosylated peptide, requires that a recombinantly-produced peptide is subjected to additional processing steps, either within the cell or after the peptide is produced by the cell, where the processing steps are performed in vitro.
- the peptide can be treated enzymatically to introduce one or more glycosyl groups onto the peptide, using a glycosyltransferase. Specifically, the glycosyltransferase covalently attaches the glycosyl group or groups to the peptide.
- Glycosyltransferases are reviewed in general in International (PCT) Patent Application No. WO03/031464 (PCT/US02/32263), which is incorporated herein by reference in its entirety.
- One such particular glycosyltransferase that has utility in the development and production of therapeutic glycopeptides is GalNAcT2.
- GalNAcT2 or N- acetyl-D-galactosamine transferase, catalyzes the transfer of GalNAc from a GalNAc donor to a GalNAc acceptor.
- Full length human GalNAcT2 enzyme is disclosed by Bennett et al. (1996, J Biol Chem. 271:17006-17012). However, the identification of useful mutants of this enzyme, having enhanced biological activity such as enhanced catalytic activity or enhanced stability, has not heretofore been reported.
- the present invention provides an isolated nucleic acid comprising a nucleic acid sequence that encodes a truncated human GalNAcT2 polypeptide.
- the truncated human GalNAcT2 polypeptide lacks all or a portion of the GalNAcT2 signal domain, or in addition lacks all or a portion the GalNAcT2 transmembrane domain, or in addition lacks all or a portion the GalNAcT2 stem domain; with the proviso that the encoded polypeptide is not a human GalNAcT2 truncation mutant polypeptide lacking amino acid residues 1-51.
- the isolated nucleic acid comprises a nucleic acid sequence having at least 90% identity with a nucleic acid selected from the group consisting of SEQ ID NO:3, SEQ ID NO:7 and SEQ ID NO:9. In another embodiment, the isolated nucleic acid comprises a nucleic acid sequence having at least 95% identity with a nucleic acid selected from the group consisting of SEQ ID NO:3, SEQ ID NO:7 and SEQ ID NO:9. In a further embodiment, the isolated nucleic acid comprises a nucleic acid sequence selected from SEQ ID NO:3, SEQ ID NO:7 and SEQ ID NO:9.
- the isolated nucleic acid is an isolated chimeric nucleic acid encoding a fusion polypeptide.
- the fusion polypeptide can include a tag polypeptide covalently linked to a truncated human GalNAcT2 polypeptide, as described herein.
- tag polypeptides include a maltose binding protein, a histidine tag, a Factor LX tag, a glutathione-S-transferase tag, a FLAG-tag, and a starch binding domain tag.
- the invention provides an isolated truncated human GalNAcT2 polypeptide, that lacks all or a portion of the GalNAcT2 signal domain, or in addition lacks all or a portion the GalNAcT2 transmembrane domain, or in addition lacks all or a portion the GalNAcT2 stem domain; with the proviso that the encoded polypeptide is not a human GalNAcT2 truncation mutant polypeptide lacking amino acid residues 1-51.
- the isolated truncated human GalNAcT2 polypeptide has at least 90% or 95% identity with a polypeptide selected from the group consisting of SEQ ID NO:4, SEQ ID NO:8 and SEQ ID NO: 10.
- isolated truncated human GalNAcT2 polypeptide comprises an amino acid sequence selected from SEQ ID NO:4, SEQ ID NO:8 and SEQ ID NO: 10.
- the isolated truncated GalNAcT2 polypeptide isolated chimeric polypeptide comprising a tag polypeptide covalently linked to the isolated truncated GalNAcT2.
- tag polypeptides include a maltose binding protein, a histidine tag, a Factor IX tag, a glutathione-S-transferase tag, a FLAG-tag, and a starch binding domain tag.
- the isolated nucleic acid encoding a truncated GalNAcT2 polypeptide can also be operably linked to a promoter/regulatory sequence, within e.g., an expression vector.
- the invention also includes host cells that comprise such expression vectors.
- Host cells can be e.g., eukaryotic or a prokaryotic cells.
- Eukaryotic cells include, e.g., mammalian cells, an insect cells, and a fungal cells. Some preferred mammalian host cells are SF9 cells, an SF9+ cells, an Sf21 cells, a HIGH FIVE cells or Drosophila Schneider S2 cells.
- Prokaryotic host cells include, e.g., E. coli cells and 5. subtilis cells.
- the host cells can be used to producing a truncated human GalNAcT2 polypeptide, by growing the recombinant host cells of under conditions suitable for expression of the truncated human GalNAcT2 polypeptide.
- sufficient truncated human GalNAcT2 polypeptide is made to allow commercial scale production of a glycoprotein or glycopeptide.
- the invention includes a method of catalyzing the transfer of a
- GalNAc moiety to an acceptor moiety comprising incubating the truncated human GalNAcT2 polypeptide with a GalNAc moiety and an acceptor moiety, wherein said polypeptide mediates the covalent linkage of said GalNAc moiety to said acceptor moiety, thereby catalyzing the transfer of a GalNAc moiety to an acceptor moiety to produce a product saccharide, or a product glycoprotein, or a product glycopeptide.
- the acceptor moiety is a granulocyte colony stimulating factor (G-CSF) protein.
- G-CSF granulocyte colony stimulating factor
- the acceptor moiety is selected from erythropoietin, human growth hormone, granulocyte colony stimulating factor, interferons alpha, -beta, and -gamma, Factor IX, follicle stimulating hormone, interleukin-2, erythropoietin, anti-TNF-alpha, and a lysosomal hydrolase.
- the polypeptide acceptor is a glycopeptide.
- the GalNAc moiety comprises a polyethylene glycol moiety.
- the product saccharide, product glycoprotein, or product glycopeptide is produced on a commercial scale.
- FIG. 1 is an image of an electrophoretic gel illustrating the PCR amplification of ppGalNAcT2 genes.
- PCR1 PCR product for ppGalNAcT2-N41R (1596 bp); PCR2, PCR product for ppGalNAcT2-N52K (1563 bp); PCR3, PCR product for ppGalNAcT2-N74G (1497 bp); PCR4, PCR product for ppGalNAcT2-N95G (1434 bp).
- Figure 2 A is a plasmid restriction map for the pCWin2MBP vector.
- Figure 2B is an image of an electrophoretic gel illustrating the fragments resulting from multiple samples of the pCWin2MBP vector digested by both BamHI and Xhol restriction enzymes.
- Figure 3 is an image of an electrophoretic gel illustrating the screening of DH5 ⁇ (pCWin2MBP-ppGalNAcT2) colonies by restriction mapping (BamHI and Xhol digestion) for plasmid purified from twelve colonies.
- Lane M bp ladder. Lanes 1-3, N41R; lanes 4-6, N52K; lanes 8-10, N74G; lanes 11-13, N95G.
- FIG 4 is an image of an electrophoretic protem gel illustrating SDS-PAGE for JM109 ( ⁇ CWin2MBP-ppGalNAcT2) whole cell lysates after IPTG induction as described elsewhere herein.
- M Pre-Stained MW Standard; Lane 13, IPTG-induced JM109
- FIG. 5 is an image of an electrophoretic protein gel illustrating SDS-PAGE for JM109 (pCWin2MBP-ppGalNAcT2) cell lysates.
- M Pre-Stained MW Standard; Lane 13, lysate from JM109 (pCWin2MBP); Lanes 1-12, lysates from colonies 1-12; Lanes 1-3, JM109 (pCWin2MBP-ppGalNAcT2N41R); Lanes 4-6, JM109 (pCWin2MBP- ppGalNAcT2N52K); Lanes 7-9, JM109 (pCWin2MBP-ppGalNAcT2N74G); Lanes 10-12, JM109 (pCWin2MBP-ppGalNAcT2N95G).
- Figure 6 is an image of an electrophoretic protein gel illustrating SDS-PAGE for inclusion bodies isolated from JM109 (pCWin2MBP- ⁇ GalNAcT2) cells.
- M Pre-Stained MW Standard; Lane 13, inclusion bodies from JM109 (pCWin2MBP); Lanes 1-12, inclusion bodies from colonies 1-12; Lanes 1-3, JM109 (pCWm2MBP-ppGalNAcT2N41R); Lanes 4-6, JM109 (pCWin2MBP-ppGalNAcT2N52K); Lanes 7-9, JM109 (pCWin2MBP- ppGalNAcT2N74G); Lanes 10-12, JM109 (pCWin2MBP-ppGalNAcT2N95G).
- Figure 7 is an image of an electrophoretic gel illustrating the protein expression pattern in lysates of cells containing human GalNAcT2 constructs.
- Lane 1 molecular weight marker
- lane 2 construct 1 culture before induction
- lane 3 construct 1 culture after induction
- lane 4 construct 2 culture before induction
- lane 5 construct 2 culture after induction
- lane 9, construct 4 culture after induction lane 10, empty.
- Figure 8 is an image of an electrophoretic protein gel illustrating the protein content of inclusion bodies from JM109 pCWir ⁇ MBP-GalNAcT2 constructs. Lane 1, MW marker; lane 2, JM109 ⁇ CWin2 MBP-GalNAcT2 construct 1 inclusion bodies; lane 3, JM109 pCWin2 MBP-GalNAcT2 construct 2 inclusion bodies.
- Figure 9 is an image of an electrophoretic protein gel illustrating the glycoPEGylation of G-CSF by ⁇ 51 GalNAcT2-MBP.
- Lane 1 glycoPEGylation in the presence of 1 mg/ml G-CSF;
- lane 2 glycoPEGylation in the presence of 0.7 mg/ml G-CSF;
- lane 3 glycoPEGylation in the presence of 0.4 mg/ml G-CSF;
- lane 4 glycoPEGylation in the presence of 0.2 mg/ml G-CSF.
- the glycoPEGylated G-CSF is visible around 60 kDa.
- Figures 10A and 10B depict a nucleic acid sequence encoding a ⁇ 40 GalNAcT2 polypeptide.
- Figures 11 A and 1 IB depict a nucleic acid sequence encoding a ⁇ 51 GalNAcT2 polypeptide.
- Figures 12A and 12B depict a nucleic acid sequence encoding a ⁇ 73 GalNAcT2 polypeptide.
- Figures 13 A and 13B depict a nucleic acid sequence encoding a ⁇ 94 GalNAcT2 polypeptide.
- Figure 14A is an image of a chromatogram illustrating the elution of ⁇ 51
- GalNAcT2-MBP that was refolded at pH 5.5 and subsequently eluted from a Q-sepharose fast flow column. Fraction numbers are indicated on the X-axis and the relative absorbance of each fraction is indicated on the Y-axis.
- Figure 14B is an image of two electrophoretic gels used to visualize the eluted fractions set forth in Figure 14A. The contents of each lane on the gel are described in the figure.
- Figure 14C is a table illustrating the relative GalNAc transferase activity of the fractions set forth in Figure 14 A.
- Figure 15 A is an image of a chromatogram illustrating the elution of ⁇ 51 GalNAcT2-MBP that was refolded at pH 6.5 and subsequently eluted from a Q-sepharose fast flow column. Fraction numbers are indicated on the X-axis and the relative absorbance of each fraction is indicated on the Y-axis.
- Figure 15B is an image of two electrophoretic gels used to visualize the eluted fractions set forth in Figure 15 A. The contents of each lane on the gel are described in the figure.
- Figure 15C is a table illustrating the relative GalNAc transferase activity of the fractions set forth in Figure 15 A.
- Figure 16 A is an image of a chromatogram illustrating the elution of ⁇ 51 GalNAcT2-MBP that was refolded at pH 8.0 and subsequently eluted from a Q-sepharose fast flow column. Fraction numbers are indicated on the X-axis and the relative absorbance of each fraction is indicated on the Y-axis.
- Figure 16B is an image of two electrophoretic gels used to visualize the eluted fractions set forth in Figure 16A. The contents of each lane on the gel are described in the figure.
- Figure 16C is a table illustrating the relative GalNAc transferase activity of the fractions set forth in Figure 16A.
- Figure 17A is an image of a chromatogram illustrating the elution of ⁇ 51 GalNAcT2-MBP that was refolded at pH 8.5 and subsequently eluted from a Q-sepharose fast flow column. Fraction numbers are indicated on the X-axis and the relative absorbance of each fraction is indicated on the Y-axis.
- Figure 17B is an image of two electrophoretic gels used to visualize the eluted fractions set forth in Figure 17 A. The contents of each lane on the gel are described in the figure.
- Figure 17C is a table illustrating the relative GalNAc transferase activity of the fractions set forth in Figure 17A.
- Figure 18 A is an image of a chromatogram illustrating the elution of ⁇ 51 GalNAcT2-MBP that was refolded at pH 8.0 and subsequently eluted from a Q-sepharose fast flow column. Fraction numbers are indicated on the X-axis and the relative absorbance of each fraction is indicated on the Y-axis.
- Figure 18B is an image of two electrophoretic gels used to visualize the eluted fractions set forth in Figure 18 A. The contents of each lane on the gel are described in the figure.
- Figure 18C is a table illustrating the relative GalNAc transferase activity of the fractions set forth in Figure 18 A.
- Figure 19A is an image of a chromatogram illustrating the elution of ⁇ 51
- GalNAcT2-MBP from a Q-sepharose fast flow column. Fraction numbers are indicated on the X-axis and the relative absorbance of each fraction is indicated on the Y-axis.
- Figure 19B is an image of two electrophoretic gels used to visualize the eluted fractions set forth in Figure 19A. The contents of each lane on the gel are described in the figure and correspond to the chromatogram of Figure 19 A.
- Figure 19C is a table illustrating the relative GalNAc transferase activity of the fractions set forth in Figure 19 A.
- Figure 20 A is an image of a chromatogram illustrating the elution of ⁇ 51 GalNAcT2-MBP from a Q-sepharose XL column, using 5 mM NaCl. Fraction numbers are indicated on the X-axis and the relative absorbance of each fraction is indicated on the Y- axis.
- Figure 20B is an image of two electrophoretic gels used to visualize the eluted fractions set forth in Figure 20A. The contents of each lane on the gel are described in the figure and correspond to the chromatogram of Figure 20 A.
- Figure 20C is a table illustrating the relative GalNAc transferase activity of the fractions set forth in Figure 20A.
- Figure 21 A is an image of a chromatogram illustrating the elution of ⁇ 51 GalNAcT2-MBP from a Q-sepharose XL column, using 50 mM NaCl. Fraction numbers are indicated on the X-axis and the relative absorbance of each fraction is indicated on the Y- axis.
- Figure 2 IB is an image of two electrophoretic gels used to visualize the eluted fractions set forth in Figure 21 A. The contents of each lane on the gel are described in the figure and correspond to the chromatogram of Figure 21 A.
- Figure 21 C is a table illustrating the relative GalNAc transferase activity of the fractions set forth in Figure 21 A.
- Figure 22 A is an image of a chromatogram illustrating the elution of ⁇ 51 GalNAcT2-MBP from a Q-sepharose XL column, using 100 mM NaCl. Fraction numbers are indicated on the X-axis and the relative absorbance of each fraction is indicated on the Y- axis.
- Figure 22B is an image of two electrophoretic gels used to visualize the eluted fractions set forth in Figure 22A. The contents of each lane on the gel are described in the figure and correspond to the chromatogram of Figure 22 A.
- Figure 22C is a table illustrating the relative GalNAc transferase activity of the fractions set forth in Figure 22A.
- Figure 23 A is an image of a chromatogram illustrating the elution of ⁇ 51
- GalNAcT2-MBP from a Q-sepharose XL column, using 200 mM NaCl. Fraction numbers are indicated on the X-axis and the relative absorbance of each fraction is indicated on the Y- axis.
- Figure 23B is an image of two electrophoretic gels used to visualize the eluted fractions set forth in Figure 23 A. The contents of each lane on the gel are described in the figure and correspond to the chromatogram of Figure 23 A.
- Figure 23 C is a table illustrating the relative GalNAc transferase activity of the fractions set forth in Figure 23 A.
- Figure 24A is an image of a chromatogram illustrating the elution of ⁇ 51 GalNAcT2-MBP from a Hydroxyapatite Type I column. Fraction numbers are indicated on the X-axis and the relative absorbance of each fraction is indicated on the Y-axis.
- Figure 24B is an image of an electrophoretic gel used to visualize the eluted fractions set forth in Figure 24A. The contents of each lane on the gel are described in the figure and correspond to the chromatogram of Figure 24A.
- Figure 24C is a table illustrating the relative GalNAc transferase activity of the fractions set forth in Figure 24A.
- Figure 25 is a graph illustrating the relative GalNAc transferase activity of various preparations of refolded ⁇ 51 GalNAcT2-MBP. The refolding conditions of each preparation is indicated on the x-axis, and the relative GalNAc transferase activity is illustrated on the Y- axis.
- Figure 26 is a graph illustrating the relative GalNAc transferase activity of various preparations of refolded ⁇ 51 GalNAcT2-MBP. The refolding conditions of each preparation is indicated on the x-axis, and the relative GalNAc transferase activity is illustrated on the Y- axis.
- Figure 27 is an image of three MALDI-TOF spectra demonstrating GalNAc transfer to GCSF mediated by ⁇ 51 GalNAcT2-MBP that has been refolded and purified according to the present invention.
- Figure 28 is an image of three MALDI-TOF spectra demonstrating GalNAc transfer to GCSF mediated by ⁇ 51 GalNAcT2-MBP that has been refolded and purified according to the present invention.
- compositions and methods of the present invention encompass truncation mutants of human GalNAcT2 polypeptides, isolated nucleic acids encoding these proteins, and methods of their use.
- GalNAcT2 polypeptides catalyze the transfer of a GalNAc from a GalNAc donor to a GalNAc acceptor.
- the glycosyltransferase GalNAcT2 is an essential reagent for glycosylation of therapeutic glycopeptides. Additionally, GalNAc T2 is an important reagent for research and development of therapeutically important glycopeptides and oligosaccharide therapeutics. GalNAcT2 enzymes are typically isolated and purified from natural sources, or from tedious and costly in vitro and recombinant sources.
- the present invention provides compositions and methods relating to simplified and more cost-effective methods of production of GalNAcT2 enzymes. In particular, the present invention provides compositions and methods relating to truncated GalNAcT2 enzymes that have improved and useful properties in comparison to their full-length enzyme counterparts.
- Truncated glycosyltransferase enzymes of the present invention are useful for in vivo and in vitro preparation of glycosylated peptides, as well as for the production of oligosaccharides containing the specific glycosyl residues that can be transferred by the truncated glycosyltransferase enzymes of the present invention. This is because it is shown for the first time herein that truncated forms of GalNAcT2 polypeptides possess biological activities comparable to, and in some instances, in excess of their full-length polypeptide counterparts. The present application also discloses that such truncation mutants not only possess biological activity, but also that the truncation mutants may have enhanced properties of solubility, stability and resistance to proteolytic degradation.
- Encoding refers to the inherent property of specific sequences of nucleotides in a nucleic acid, such as a gene, a cDNA, or an mRNA, to serve as templates for synthesis of other polymers and macromolecules in biological processes having either a defined sequence of nucleotides (i.e., rRNA, tRNA and mRNA) or a defined sequence of amino acids and the biological properties resulting therefrom.
- a gene encodes a protein if transcription and translation of mRNA corresponding to that gene produces the protein in a cell or other biological system.
- Both the coding strand the nucleotide sequence of which is identical to the mRNA sequence and is usually provided in sequence listings, and the non-coding strand, used as the template for transcription of a gene or cDNA, can be referred to as encoding the protein or other product of that gene or cDNA.
- a "coding region" of a gene consists of the nucleotide residues of the coding strand of the gene and the nucleotides of the non-coding strand of the gene which are homologous with or complementary to, respectively, the coding region of an mRNA molecule which is produced by transcription of the gene.
- a "coding region" of an mRNA molecule also consists of the nucleotide residues of the mRNA molecule which are matched with an anticodon region of a transfer RNA molecule during translation of the mRNA molecule or which encode a stop codon.
- the coding region may thus include nucleotide residues corresponding to amino acid residues which are not present in the mature protein encoded by the mRNA molecule (e.g., amino acid residues in a protein export signal sequence).
- An "affinity tag” is a peptide or polypeptide that may be genetically or chemically fused to a second polypeptide for the purposes of purification, isolation, targeting, trafficking, or identification of the second polypeptide.
- the "genetic" attachment of an affinity tag to a second protein may be effected by cloning a nucleic acid encoding the affinity tag adjacent to a nucleic acid encoding a second protein in a nucleic acid vector.
- glycosyltransferase refers to any enzyme/protein that has the ability to transfer a donor sugar to an acceptor moiety.
- a "sugar nucleotide-generating enzyme” is an enzyme that has the ability to produce a sugar nucleotide.
- Sugar nucleotides are known in the art, and include, but are not limited to, such moieties as UDP-Gal, UDP-GalNAc, and CMP-NAN.
- isolated nucleic acid refers to a nucleic acid segment or fragment which has been separated from sequences which flank it in a naturally occurring state, e.g., a DNA fragment which has been removed from the sequences which are normally adjacent to the fragment, e.g., the sequences adjacent to the fragment in a genome in which it naturally occurs.
- the term also applies to nucleic acids which have been substantially purified from other components which naturally accompany the nucleic acid, e.g., RNA or DNA or proteins, which naturally accompany it in the cell.
- the term therefore includes, for example, a recombinant DNA which is incorporated into a vector, into an autonomously replicating plasmid or virus, or into the genomic DNA of a prokaryote or eukaryote, or which exists as a separate molecule (e.g, as a cDNA or a genomic or cDNA fragment produced by PCR or restriction enzyme digestion) independent of other sequences. It also includes a recombinant DNA which is part of a hybrid gene encoding additional polypeptide sequence.
- A refers to adenosine
- C refers to cytidine
- G refers to guanosine
- T refers to thymidine
- U refers to uridine.
- a "polynucleotide” means a single strand or parallel and anti-parallel strands of a nucleic acid.
- a polynucleotide may be either a single-stranded or a double-stranded nucleic acid.
- nucleic acid typically refers to large polynucleotides. However, the terms “nucleic acid” and “polynucleotide” are used interchangeably herein.
- oligonucleotide typically refers to short polynucleotides, generally no greater than about 50 nucleotides. It will be understood that when a nucleotide sequence is represented by a DNA sequence (i.e., A, T, G, C), this also includes an RNA sequence (i.e., A, U, G, C) in which "U" replaces "T.”
- nucleic acid sequences the left- hand end of a single-stranded nucleic acid sequence is the 5' end; the left-hand direction of a double-stranded nucleic acid sequence is referred to as the 5'-direction.
- a first defined nucleic acid sequence is said to be "immediately adjacent to" a second defined nucleic acid sequence when, for example, the last nucleotide of the first nucleic acid sequence is chemically bonded to the first nucleotide of the second nucleic acid sequence through a phosphodiester bond.
- a first defined nucleic acid sequence is also said to be "immediately adjacent to" a second defined nucleic acid sequence when, for example, the first nucleotide of the first nucleic acid sequence is chemically bonded to the last nucleotide of the second nucleic acid sequence through a phosphodiester bond.
- a first defined polypeptide sequence is said to be "immediately adjacent to" a second defined polypeptide sequence when, for example, the last amino acid of the first polypeptide sequence is chemically bonded to the first amino acid of the second polypeptide sequence through a peptide bond.
- a first defined polypeptide sequence is said to be "immediately adjacent to" a second defined polypeptide sequence when, for example, the first amino acid of the first polypeptide sequence is chemically bonded to the last amino acid of the second polypeptide sequence through a peptide bond.
- the direction of 5' to 3' addition of nucleotides to nascent RNA transcripts is referred to as the transcription direction.
- the DNA strand having the same sequence as an mRNA is referred to as the "coding strand”; sequences on the DNA strand which are located 5' to a reference point on the DNA are referred to as “upstream sequences”; sequences on the DNA strand which are 3' to a reference point on the DNA are referred to as "downstream sequences.”
- nucleotide sequence encoding an amino acid sequence includes all nucleotide sequences that are degenerate versions of each other and that encode the same amino acid sequence. Nucleotide sequences that encode proteins and RNA may include introns.
- Homology refers to nucleotide sequence similarity between two regions of the same nucleic acid strand or between regions of two different nucleic acid strands. When a nucleotide residue position in both regions is occupied by the same nucleotide residue, then the regions are homologous at that position. A first region is homologous to a second region if at least one nucleotide residue position of each region is occupied by the same residue. Homology between two regions is expressed in terms of the proportion of nucleotide residue positions of the two regions that are occupied by the same nucleotide residue.
- a region having the nucleotide sequence 5'- ATTGCC-3' and a region having the nucleotide sequence 5'-TATGGC-3' share 50% homology.
- the first region comprises a first portion and the second region comprises a second portion, whereby, at least about 50%, and preferably at least about 75%, at least about 90%, or at least about 95% of the nucleotide residue positionss of each of the portions are occupied by the same nucleotide residue. More preferably, all nucleotide residue positions of each of the portions are occupied by the same nucleotide residue.
- percent identity is used synonymously with “homology.”
- the determination of percent identity between two nucleotide or amino acid sequences can be accomplished using a mathematical algorithm.
- a mathematical algorithm useful for comparing two sequences is the algorithm of Karlin and Altschul (1990, Proc. Natl. Acad. Sci. USA 87:2264-2268), modified as in Karlin and Altschul (1993, Proc. Natl. Acad. Sci. USA 90:5873-5877). This algorithm is incorporated into the NBLAST and XBLAST programs of Altschul et al. (1990, J. Mol. Biol.
- BLAST protein searches can be performed with the XBLAST program (designated “blastn” at the NCBI web site) or the NCBI “blastp” program, using the following parameters: expectation value 10.0, BLOSUM62 scoring matrix to obtain amino acid sequences homologous to a protein molecule described herein.
- Gapped BLAST can be utilized as described in Altschul et al. (1997, Nucleic Acids Res. 25:3389-3402).
- PSI-Blast or PHI-Blast can be used to perform an iterated search which detects distant relationships between molecules (id.) and relationships between molecules which share a common pattern.
- the default parameters of the respective programs e.g., XBLAST and NBLAST
- the default parameters of the respective programs can be used as available on the website of the National Center for Biotechnology Information of the National Library of Medicine at the National Institutes of Health.
- the percent identity between two sequences can be determined using techniques similar to those described above, with or without allowing gaps. In calculating percent identity, typically exact matches are counted.
- Polypeptide refers to a polymer composed of amino acid residues, related naturally occurring structural variants, and synthetic non-naturally occurring analogs thereof linked via peptide bonds, related naturally occurring structural variants, and synthetic non- naturally occurring analogs thereof. Synthetic polypeptides can be synthesized, for example, using an automated polypeptide synthesizer. A "polypeptide,” as the term is used herein, therefore refers to any size polymer of amino acid residues, provided that the polymer contains at least two amino acid residues.
- protein typically refers to large peptides, also referred to herein as “polypeptides.”
- peptide typically refers to short polypeptides.
- peptide may refer to an amino acid polymer of three amino acids, as well as an amino acid polymer of several hundred amino acids.
- amino acids are represented by the full name thereof, by the three letter code corresponding thereto, or by the one-letter code corresponding thereto, as indicated in the following table:
- a "therapeutic peptide” as the term is used herein refers to any peptide that is useful to treat a disease state or to improve the overall health of a living organism.
- a therapeutic peptide may effect such changes in a living organism when administered alone, or when used to improve the therapeutic capacity of another substance.
- the term “therapeutic peptide” is used interchangeably herein with the terms “therapeutic polypeptide” and “therapeutic protein.”
- a "reagent peptide” as the term is used herein refers to any peptide that is useful in food biochemistry, bioremediation, production of small molecule therapeutics, and even in the production of therapeutic peptides.
- reagent peptides are enzymes capable of catalyzing a reaction to produce a product useful in any of the aforementioned areas.
- the term “reagent peptide” is used interchangeably herein with the terms “reagent polypeptide” and "reagent protein. "
- glycopeptide refers to a peptide having at least one carbohydrate moiety covalently linked thereto. It will be understood that a glycopeptide may be a "therapeutic glycopeptide,” as described above.
- glycopeptide is used interchangeably herein with the terms “glycopolypeptide” and “glycoprotein.”
- a "vector” is a composition of matter which comprises an isolated nucleic acid and which can be used to deliver the isolated nucleic acid to the interior of a cell.
- vectors are known in the art including, but not limited to, linear nucleic acids, nucleic acids associated with ionic or amphiphilic compounds, plasmids, and viruses.
- the te ⁇ n "vector” includes an autonomously replicating plasmid or a virus.
- the term should also be construed to include non-plasmid and non- viral compounds which facilitate transfer of nucleic acid into cells, such as, for example, polylysine compounds, liposomes, and the like.
- viral vectors include, but are not limited to, adenoviral vectors, adeno-associated virus vectors, retroviral vectors, and the like.
- Expression vector refers to a vector comprising a recombinant nucleic acid comprising expression control sequences operatively linked to a nucleotide sequence to be expressed.
- An expression vector comprises sufficient cis-acting elements for expression; other elements for expression can be supplied by the host cell or in an in vitro expression system.
- Expression vectors include all those known in the art, such as cosmids, plasmids (e.g., naked or contained in liposomes) and viruses that incorporate the recombinant nucleic acid.
- a "multiple cloning site" as the term is used herein is a region of a nucleic acid vector that contains more than one sequence of nucleotides that is recognized by at least one restriction enzyme.
- an "antibiotic resistance marker” as the term is used herein refers to a sequence of nucleotides that encodes a protein which, when expressed in a living cell, confers to that cell the ability to live and grow in the presence of an antibiotic.
- GalNAcT2 refers to N-acetyl-D-galactosamine transferase 2.
- a "truncated" form of a peptide refers to a peptide that is lacking one or more amino acid residues as compared to the full-length amino acid sequence of the peptide.
- the peptide "NH2-Ala-Glu-Lys-Leu-COOH” is an N-terminally truncated form of the full-length peptide "NH2-Gly-Ala-Glu-Lys-Leu-COOH.”
- the terms "truncated form” and “truncation mutant” are used interchangeably herein.
- a truncated peptide is a GalNAcT2 polypeptide comprising an active domain, a stem domain, a transmembrane domain, and a signal domain, wherein the signal domain is lacking a single N-terminal amino acid residue as compared to the full length GalNAcT2.
- saccharides refers in general to any carbohydrate, a chemical entity with the most basic structure of (CH 2 O) n . Saccharides vary in complexity, and may also include nucleic acid, amino acid, or virtually any other chemical moiety existing in biological systems.
- Olet al. refers to a molecule consisting of several units of carbohydrates of defined identity. Typically, saccharide sequences between 2-20 units may be referred to as oligosaccharides.
- Polysaccharide refers to a molecule consisting of many units of carbohydrates of defined identity. However, any saccharide of two or more units may correctly be considered a polysaccharide.
- a saccharide "donor” is a moiety that can provide a saccharide to a glycosyltransferase so that the glycosyltransferase may transfer the saccharide to a saccharide acceptor.
- a GalNAc donor may be UDP-GalNAc.
- a saccharide "acceptor” is a moiety that can accept a saccharide from a saccharide donor.
- a glycosyltransferase can covalently couple a saccharide to a saccharide acceptor.
- G-CSF may be a GalNAc acceptor, and a GalNAc moiety may be covalently coupled to a GalNAc acceptor by way of a GalNAc- transferase.
- a saccharide acceptor is a protein or peptide comprising an O glycosylation site.
- saccharide acceptors include, e.g., erythropoietin, human growth hormone, granulocyte colony stimulating factor, interferons alpha, -beta, and -gamma, Factor IX, follicle stimulating hormone, interleukin-2, erythropoietin, anti-TNF-alpha, and a lysosomal hydrolase
- An oligosaccharide with a "defined size” is one which consists of an identifiable number of monosaccharide units.
- an oligosaccharide consisting of 10 monosaccharide units is one which may consist of 10 identical monosaccharide units or 5 monosaccharide units of a first identity and 5 monosaccharide units of a second identity.
- an oligosaccharide of defined size that consists of monosaccharide units of heterogeneous identity may have the monosaccharide units in any order from beginning to end of the oligosaccharide.
- An oligosaccharide of "random size" is one which may be synthesized using methods that do not provide oligosaccharide products of defined size.
- a method of oligosaccharide synthesis may provide oligosaccharides that range from two monosaccharide units to twenty-two saccharide units, including any or all lengths in between.
- Communication scale refers to gram scale production of a product saccharide, or glycoprotein, or glycopeptide in a single reaction. In preferred embodiments, commercial scale refers to production of greater than about 50, 75, 80, 90 or 100, 125, 150, 175, or 200 grams.
- sialic acid refers to any member of a family of nine-carbon carboxylated sugars.
- the most common member of the sialic acid family is N-acetyl-neuraminic acid (2- keto-5-acetamido-3,5-dideoxy-D-glycero-D-galactononulopyranos-l-onic acid (often abbreviated as Neu5Ac, NeuAc, or NANA).
- a second member of the family is N-glycolyl- neuraminic acid (Neu5Gc or NeuGc), in which the N-acetyl group of NeuAc is hydroxylated.
- a third sialic acid family member is 2-keto-3-deoxy-nonulosonic acid (KDN) (Nadano et al. (1986) J. Biol. Chem. 261: 11550-11557; Kanamori et al, J. Biol. Chem. 265: 21811-21819 (1990)). Also included are 9-substituted sialic acids such as a 9-O-C ⁇ -C ⁇ acyl-Neu5Ac like 9-O-lactyl-Neu5Ac or 9-O-acetyl-Neu5Ac, 9-deoxy-9-fluoro-Neu5Ac and 9-azido-9-deoxy- Neu5Ac.
- KDN 2-keto-3-deoxy-nonulosonic acid
- 9-substituted sialic acids such as a 9-O-C ⁇ -C ⁇ acyl-Neu5Ac like 9-O-lactyl-Neu5Ac or 9-O-acety
- sialic acid family see, e.g., Varki, Glycobiology 2: 25-40 (1992); Sialic Acids: Chemistry, Metabolism and Function, R. Schauer, Ed. (Springer-Verlag, New York (1992)).
- the synthesis and use of sialic acid compounds in a sialylation procedure is disclosed in international application WO 92/16640, published October 1, 1992.
- a "method of remodeling a protein, a peptide, a glycoprotein, or a glycopeptide” as used herein, refers to addition of a sugar residue to a protein, a peptide, a glycoprotein, or a glycopeptide using a glycosyltransferase.
- the sugar residue is covalently attached to a PEG molecule.
- an "unpaired cysteine residue” as used herein, refers to a cysteine residue, which in a correctly folded protein (i.e., a protein with biological activity), does not form a disulfide bind with another cysteine residue.
- an "insoluble glycosyltransferase” refers to a glycosyltransferase that is expressed in bacterial inclusion bodies. Insoluble glycosyltransferases are typically solubilized or denatured using e.g., detergents or chaotropic agents or some combination. "Refolding” refers to a process of restoring the structure of a biologically active glycosyltransferase to a glycosyltransferase that has been solubilized or denatured. Thus, a refolding buffer, refers to a buffer that enhances or accelerates refolding of a glycosyltransferase.
- a "redox couple” refers to mixtures of reduced and oxidized thiol reagents and include reduced and oxidized glutathione (GSH/GSSG), cysteine/cystine, cysteamine/cystamine, DTT/GSSG, and DTE/GSSG. (See, e.g., Clark, Cur. Op. Biotech. 12:202-207 (2001)).
- contacting is used herein interchangeably with the following: combined with, added to, mixed with, passed over, incubated with, flowed over, etc.
- PEG refers to poly(ethylene glycol).
- PEG is an exemplary polymer that has been conjugated to peptides.
- the use of PEG to derivatize peptide therapeutics has been demonstrated to reduce the immunogenicity of the peptides and prolong the clearance time from the circulation.
- U.S. Pat. No. 4,179,337 (Davis et al.) concerns non- immunogenic peptides, such as enzymes and peptide hormones coupled to polyethylene glycol (PEG) or polypropylene glycol. Between 10 and 100 moles of polymer are used per mole peptide and at least 15% of the physiological activity is maintained.
- the term "specific activity" as used herein refers to the catalytic activity of an enzyme, e.g., a recombinant glycosyltransferase fusion protein of the present invention, and may be expressed in activity units.
- one activity unit catalyzes the formation of 1 ⁇ mol of product per minute at a given temperature (e.g., at 37°C) and pH value (e.g., at pH 7.5).
- 10 units of an enzyme is a catalytic amount of that enzyme where 10 ⁇ mol of substrate are converted to 10 ⁇ mol of product in one minute at a temperature of, e.g., 37 °C and a pH value of, e.g., 7.5.
- N-linked oligosaccharides are those oligosaccharides that are linked to a peptide backbone through asparagine, by way of an asparagine-N-acetylglucosamine linkage. N- linked oligosaccharides are also called “N-glycans.” All N-linked oligosaccharides have a common pentasaccharide core of Man 3 GlcNAc 2 . They differ in the presence of, and in the number of branches (also called antennae) of peripheral sugars such as N-acetylglucosamine, galactose, N-acetylgalactosamine, fucose and sialic acid. Optionally, this structure may also contain a core fucose molecule and/or a xylose molecule.
- O-linked oligosaccharides are those oligosaccharides that are linked to a peptide backbone through threonine, serine, hydroxyproline, tyrosine, or other hydroxy-containing amino acids.
- substantially in the above definitions of "substantially uniform” generally means at least about 60%, at least about 70%, at least about 80%, or more preferably at least about 90%, and still more preferably at least about 95% of the acceptor substrates for a particular glycosyltransferase are glycosylated.
- a "fusion protein” refers to a protein comprising amino acid sequences that are in addition to, in place of, less than, and/or different from the amino acid sequences encoding the original or native full-length protein or subsequences thereof.
- a "stem region" with reference to glycosyltransferases refers to a protein domain, or a subsequence thereof, which in the native glycosyltransferases is located adjacent to the trans-membrane domain, and has been reported to function as a retention signal to maintain the glycosyltransferase in the Golgi apparatus and as a site of proteolytic cleavage.
- Stem regions generally start with the first hydrophilic amino acid following the hydrophobic transmembrane domain and end at the catalytic domain, or in some cases the first cysteine residue following the transmembrane domain.
- Exemplary stem regions include, but is not limited to, the stem region of eukaryotic ST6GalNAcI, amino acid residues from about 30 to about 207 (see e.g., the murine enzyme), amino acids 35-278 for the h uman enzyme or amino acids 37-253 for the chicken enzyme; the stem region of mammalian GalNAcT2, amino acid residues from about 71 to about 129 (see e.g., the rat enzyme).
- a "catalytic domain” refers to a protein domain, or a subsequence thereof, that catalyzes an enzymatic reaction performed by the enzyme.
- a catalytic domain of a sialyltransferase will include a subsequence of the sialyltransferase sufficient to transfer a sialic acid residue from a donor to an acceptor saccharide.
- a catalytic domain can include an entire enzyme, a subsequence thereof, or can include additional amino acid sequences that are not attached to the enzyme, or a subsequence thereof, as found in nature.
- isolated refers to material that is substantially or essentially free from components which interfere with the activity of an enzyme.
- a saccharide, protein, or nucleic acid of the invention refers to material that is substantially or essentially free from components which normally accompany the material as found in its native state.
- an isolated saccharide, protein, or nucleic acid of the invention is at least about 80% pure, usually at least about 90%, and preferably at least about 95% pure as measured by band intensity on a silver stained gel or other method for determining purity. Purity or homogeneity can be indicated by a number of means well known in the art.
- a protein or nucleic acid in a sample can be resolved by polyacrylamide gel electrophoresis, and then the protein or nucleic acid can be visualized by staining.
- high resolution of the protein or nucleic acid may be desirable and HPLC or a similar means for purification, for example, may be utilized.
- GalNAcT2 nucleic acids encode polypeptides that have a domain structure similar to other glycosyltransferases, including an N-terminal signal domain, a transmembrane domain, a stem domain, and an active domain, wherein the active domain may comprise the majority of the amino acid sequence of such polypeptides.
- domain structure(s) extraneous to the active domain of recombinant GalNAcT2 polypeptides may have a negative effect on the solubility, stability and activity of the polypeptide in an aqueous or in vitro environment.
- the presence of a hydrophobic transmembrane domain on a recombinant GalNAcT2 polypeptide used in an in vitro reaction mixture may render the polypeptide less soluble than a recombinant GalNAcT2 polypeptide without a hydryophobic transmembrane domain, and further, may even decrease the enzymatic activity of the polypeptide by affecting or destabilizing the folded structure.
- GalNAcT2 nucleic acids that encode GalNAcT2 that is shorter than full-length GalNAcT2, for the purpose of enhancing the activity, stability and/or utility of GalNAcT2 polypeptides.
- the present invention provides such modified forms of GalNAcT2. More particularly, the present invention provides isolated nucleic acids encoding such truncated polypeptides.
- Nucleic acids of the present invention encode truncated forms of GalNacT2 polypeptides, as described in greater detail elsewhere herein.
- a truncated GalNAcT2 polypeptide encoded by a nucleic acid of the present invention also referred to herein as a "truncation mutant,” may be truncated in various ways, as would be understood by the skilled artisan.
- Examples of truncated polypeptides encoded by a nucleic acid of the present invention include, but are not limited to, a polypeptide lacking a single N-terminal residue, a polypeptide lacking a single C-terminal residue, a polypeptide lacking both an single N- terminal residue and a single C-terminal residue, a polypeptide lacking a contiguous sequence of residues from the N-terminus, a polypeptide lacking a contiguous sequence of residues from the C-terminus, and any combinations thereof.
- truncations of nucleic acids encoding GalNAcT2 polypeptides may be made for numerous reasons.
- a truncation may be made in order to remove part or all of the nucleic acid sequence encoding the signal peptide domain of an GalNAcT2.
- a truncation may be made in order to remove part or all of a nucleic acid sequence encoding a transmembrane domain of an GalNAcT2.
- removal of a part or all of a nucleic acid sequence encoding a transmembrane domain may increase the solubility or stability of the encoded GalNAcT2 polypeptide and/or may increase the level of expression of the encoded polypeptide.
- a truncation may be made in order to remove part or all of a nucleic acid sequence encoding a stem domain of an GalNAcT2.
- removal of a part or all of a nucleic acid sequence encoding a stem domain may increase the solubility or stability of the encoded GalNAcT2 polypeptide and/or may increase the level of expression of the encoded polypeptide.
- the nucleic acid residue at which a truncation is made may be a highly-conserved residue, hi another aspect of the invention, the nucleic acid residue at which a truncation is made may be selected such that the encoded polypeptide has a new N-terminal amino acid residue that will aid in the purification of the expressed polypeptide. In yet another aspect, the nucleic acid residue at which a truncation is made may be selected such that the encoded truncated polypeptide does not contain a specific secondary and/or tertiary structure.
- the present invention features nucleic acids encoding smaller than full-length GalNAcT2. That is, the present invention features a nucleic acid encoding a truncated GalNAcT2 polypeptide, provided the polypeptide expressed by the nucleic acid retains the biological activity of the full-length protein.
- a truncated polypeptide is a human truncated GalNAcT2 polypeptide.
- a nucleic acid encoding a full-length human GalNAcT2 may contain a nucleic acid sequence encoding one or more identifyable polypeptide domains in addition to the "active domain," the domain primarily responsible for the catalytic activity, of GalNAcT2. This is because it is known in that art that a full-length GalNAcT2 polypeptide, and in particular, a full-length human GalNAcT2 polypeptide, contains a signal domain, a transmembrane domain, and a stem domain, in addition to an active domain.
- a nucleic acid encoding a full-length human GalNAcT2 may encode a polypeptide that has a signal domain at the amino-terminus of the polypeptide, followed by a transmembrane domain immediately adjacent to the signal domain, followed by a stem domain that is immediately adjacent to the transmembrane domain, followed by an active domain that extends to the carboxy-terminus of the polypeptide and is located immediately adjacent to the stem domain.
- an isolated nucleic acid of the invention may encode a truncated human GalNAcT2 polypeptide, wherein the truncated human GalNAcT2 polypeptide is lacking all or a portion of the GalNAcT2 signal domain.
- an isolated nucleic acid of the invention may encode a truncated human
- a nucleic acid of the invention may encode a truncated human GalNAcT2 polypeptide, wherein the truncated human GalNAcT2 polypeptide is lacking the GalNAcT2 signal domain, the GalNAcT2 transmembrane domain and all or a portion the GalNAcT2 stem domain.
- the "biological activity of GalNAcT2" is the ability to transfer a GalNAc moiety from a GalNAc donor to an acceptor molecule.
- Full-length human GalNAcT2 the sequence of which is set forth in SEQ ID NO:l, exhibits such activity.
- the "biological activity of a GalNAcT2 truncated polypeptide” is similarly the ability to transfer a GalNAc moiety from a GalNAc donor to an acceptor molecule. That is, a truncated GalNAcT2 polypeptide of the present invention can catalyze the same glycosyltransfer reaction as the full-length GalNAcT2.
- a truncated human GalNAcT2 polypeptide encoded by a GalNAcT2 nucleic acid of the invention has the ability to transfer a GalNAc moiety from a UDP-GalNAc donor to a granulocyte-colony stimulating factor (G-CSF) acceptor, wherein such a transfer results in the O-linked covalent coupling of a GalNAc moiety to a threonine residue of G-CSF.
- G-CSF granulocyte-colony stimulating factor
- GalNAcT2 is included in the present invention provided that the truncated GalNAcT2 has GalNAcT2 biological activity.
- the methods and compositions of the invention should not be construed to be limited solely to a nucleic acid comprising a GalNAcT2 truncation mutant as disclosed herein, but rather, should be construed to encompass any nucleic acid encoding a GalNAc T2 truncated mutant, prepared in accordance with the disclosure herein, either known or unknown, which is capable of catalyzing transfer of a GalNAc to a GalNAc acceptor. Modified nucleic acid sequences, i.e.
- nucleic acid sequences having sequences that differ from the nucleic acid sequences encoding the naturally-occurring proteins are also encompassed by methods and compositions of the invention, so long as the modified nucleic acid still encodes a truncated protein having the biological activity of catalyzing the transfer of a GalNAc to a GalNAc acceptor, for example.
- modified nucleic acid sequences include modifications caused by point mutations, modifications due to the degeneracy of the genetic code or naturally occurring allelic variants, and further modifications that have been introduced by genetic engineering, i.e., by the hand of man.
- nucleic acid also specifically includes nucleic acids composed of bases other than the five biologically occurring bases (adenine, guanine, thymine, cytosine and uracil).
- the present invention features an isolated nucleic acid comprising a nucleic acid sequence that is at least about 90%, 95%, 97%, 98%, or 99% identical to a nucleic acid sequence set forth in any one of SEQ ID NO:3, SEQ ID NO:7 or SEQ ID NO:9.
- the present invention also features an isolated nucleic acid sequence comprising any one of the sequences set forth in SEQ ID NO:3, SEQ ID NO: 7 or SEQ ID NO: 9, wherein the isolated nucleic acid encodes a truncated GalNAcT2 polypeptide.
- the present invention also encompasses isolated nucleic acid molecules encoding a truncated GalNAcT2 polypeptide that contains changes in amino acid residues that are not essential for activity.
- Such polypeptides encoded by an isolated nucleic acid of the invention differ in amino acid sequence from any one of the sequences set forth in SEQ ID NO:4, SEQ ID NO: 8 or SEQ ID NO: 10, yet retain the biological activity of GalNAcT2.
- an isolated nucleic acid of the invention may include a nucleotide sequence encoding a polypeptide having an amino acid sequence that is at least about 90%, 95%, 97%, 98%, or 99% identical to the amino acid sequence of SEQ ID NO:4.
- an isolated nucleic acid of the invention may include a nucleotide sequence encoding a polypeptide that has an amino acid sequence at least about 90%, 95%, 97%o, 98%, or 99% identical to an amino acid sequence set forth in any one of SEQ ID NO:8 or SEQ ID NO:10.
- the determination of percent identity between two nucleotide or amino acid sequences can be accomplished using a mathematical algorithm.
- a mathematical algorithm useful for comparing two sequences is the algorithm of Karlin and Altschul (1990, Proc. Natl. Acad. Sci. USA 87:2264-2268), modified as in Karlin and Altschul (1993, Proc. Natl. Acad. Sci.
- NBLAST and XBLAST programs of Altschul, et al. (1990, J. Mol. Biol. 215:403- 410), and can be accessed, for example at the National Center for Biotechnology Information (NCBI) world wide web site.
- BLAST protein searches can be performed with the XBLAST program (designated “blastn” at the NCBI web site) or the NCBI “blastp” program, using the following parameters: expectation value 10.0, BLOSUM62 scoring matrix to obtain amino acid sequences homologous to a protein molecule described herein.
- Gapped BLAST can be utilized as described in Altschul et al. (1997, Nucleic Acids Res. 25:3389-3402).
- PSI-Blast or PHI- Blast can be used to perform an iterated search which detects distant relationships between molecules and relationships between molecules which share a common pattern.
- the default parameters of the respective programs can be used. See, generally, the internet website for the National Center for Biotechnology Information, which is maintained by the National Library of Medicine and the National Institutes of Health.
- a nucleic acid useful in the methods and compositions of the present invention and encoding a truncated GalNAcT2 polypeptide may have at least one nucleotide inserted into the nucleic acid sequence of such a truncated mutant.
- an additional nucleic acid encoding a truncated GalNAcT2 polypeptide may have at least one nucleotide deleted from the nucleic acid sequence.
- a GalNAcT2 nucleic acid encoding a truncated mutant and useful in the invention may have both a nucleotide insertion and a nucleotide deletion present in a single nucleic acid sequence encoding the truncated polypeptide.
- nucleic acid insertions and/or deletions may be designed into the gene for numerous reasons, including, but not limited to modification of nucleic acid stability, modification of nucleic acid expression levels, modification of expressed polypeptide stability or half-life, modification of expressed polypeptide activity, modification of expressed polypeptide properties and characteristics, and changes in glycosylation pattern. All such modifications to the nucleotide sequences encoding such proteins are encompassed by the present invention.
- nucleic acid encompassed by methods and compositions of the invention may be native or synthesized nucleic acid.
- the nucleic acid may be DNA or RNA and may exist in a double-stranded, single-stranded or partially double-stranded form. Furthermore, the nucleic acid may be found as part of a virus or other macromolecule. See, e.g., Fasbender et al, 1996, J. Biol. Chem. 272:6479-89.
- the invention includes an isolated nucleic acid encoding a truncated GalNAcT2 polypeptide operably linked to a nucleic acid comprising a promoter/regulatory sequence such that the nucleic acid is preferably capable of directing expression of the polypeptide encoded by the nucleic acid.
- the invention encompasses expression vectors and methods for the introduction of exogenous DNA into cells with concomitant expression of the exogenous DNA in those cells, as described, for example, in Sambrook et al. (Third Edition, 2001, Molecular Cloning: A Laboratory Manual, Cold Spring Harbor Laboratory, New York), and in Ausubel et al. (1997, Current Protocols in Molecular Biology, John Wiley & Sons, New York).
- Expression of a truncated GalNAcT2 polypeptide in a cell may be accomplished by generating a plasmid, viral, or other type of vector comprising a nucleic acid encoding the appropriate nucleic acid, wherein the nucleic acid is operably linked to a promoter/regulatory sequence which serves to drive expression of the encoded polypeptide, with or without tag, in cells in which the vector is introduced.
- promoters which are well known in the art which are induced in response to inducing agents such as metals, glucocorticoids, and the like, are also contemplated in the invention.
- the invention includes the use of any promoter/regulatory sequence, which is either known or unknown, and which is capable of driving expression of the truncated GalNAcT2 polypeptide operably linked thereto.
- a nucleic acid encoding a truncated GalNAcT2 polypeptide may be fused to one or more additional nucleic acids encoding a functional polypeptide.
- an affinity tag coding sequence may be inserted into a nucleic acid vector adjacent to, upstream from, or downstream from a truncated GalNAcT2 polypeptide coding sequence.
- an affinity tag will typically be inserted into a multiple cloning site in frame with the truncated GalNAcT2 polypeptide.
- an affinity tag coding sequence can be used to produce a recombinant fusion protein by concomitantly expressing the affinity tag and truncated GalNAcT2 polypeptide. The expressed fusion protein can then be isolated, purified, or identified by means of the affinity tag.
- Affinity tags useful in the present invention include, but are not limited to, a maltose binding protein, a histidine tag, a Factor IX tag, a glutathione-S-transferase tag, a FLAG-tag, and a starch binding domain tag.
- Other tags are well known in the art, and the use of such tags in the present invention would be readily understood by the skilled artisan.
- a vector comprising a truncated GalNAcT2 polypeptide of the present invention may be used to express the truncated polypeptide as either a non-fusion or as a fusion protein.
- Selection of any particular plasmid vector or other DNA vector is not a limiting factor in this invention and a wide plethora of vectors are well-known in the art. Further, it is well within the skill of the artisan to choose particular promoter/regulatory sequences and operably link those promoter/regulatory sequences to a DNA sequence encoding a truncated GalNAcT2 polypeptide.
- a vector useful in one embodiment of the present invention is based on the pcWori+ vector (Muchmore et al., 1987, Meth. Enzymol. 177:44-73).
- the invention thus includes a vector comprising an isolated nucleic acid encoding a truncated GalNAcT2 polypeptide.
- a nucleic acid encoding a truncated GalNAcT2 polypeptide.
- the incorporation of a nucleic acid into a vector and the choice of vectors is well-known in the art as described in, for example, Sambrook et al. (Third Edition, 2001, Molecular Cloning: A Laboratory Manual, Cold Spring Harbor Laboratory, New York), and in Ausubel et al. (1997, Current Protocols in Molecular Biology, John Wiley & Sons, New York).
- an isolated nucleic acid encoding a truncated GalNAcT2 polypeptide is integrated into the genome of a host cell in conjunction with a nucleic acid encoding a truncated GalNAcT2 polypeptide.
- a cell is transiently transfected with an isolated nucleic acid encoding a truncated GalNAcT2 polypeptide.
- a cell is stably transfected with an isolated nucleic acid encoding a truncated GalNAcT2 polypeptide.
- a nucleic acid encoding a truncated GalNAcT2 polypeptide may be purified by any suitable means, as are well known in the art.
- the nucleic acids can be purified by reverse phase or ion exchange HPLC, size exclusion chromatography or gel electrophoresis.
- the method of purification will depend in part on the size of the DNA to be purified.
- the present invention also features a recombinant bacterial host cell comprising , inter alia, a nucleic acid vector as described elsewhere herein.
- the recombinant cell is transformed with a vector of the present invention.
- the transformed vector need not be integrated into the cell genome nor does it need to be expressed in the cell. However, the transformed vector will be capable of being expressed in the cell.
- E. coli is used for transformation of a vector of the present invention and expression of protein therefrom.
- a K-12 strain of E. coli is useful for expression of protein from a vector of the present invention.
- Strains of E. coli useful in the present invention include, but are not limited to, JM83, JM101, JM103, JM109, W3110, chil776, and JA221.
- a host cell useful in the present invention will be capable of growth and culture on a small scale, medium scale, or a large scale.
- a host cell of the invention is useful for testing the expression of a protein from a vector of the invention equally as much as it is useful for large scale production of a reagent or therapeutic protein product.
- Techniques useful in culturing host cells and expressing protein from a vector contained therein are well known in the art and will therefore not be listed herein.
- a host cell useful in methods of the present invention may be prepared according to various methods, as would be understood by the skilled artisan when armend with the disclosure set forth herein.
- a host cell of the present invention may be transformed with a vector of the present invention to produce a transformed host cell of the invention. Transformation, as known to the skilled artisan, includes the process of inserting a nucleic acid vector into a host cell, such that the host cell containing the nucleic acid vector remains viable.
- Such transformation of nucleic acid into a bacterial cell is useful for purposes including, but not limited to, creation of a stably-transformed host cell, making a biological deposit, propagating the vector-containing host cell, propagating the vector- containing host cell for the production and isolation of additional vector, expression of target protein encoded by vector, and the like.
- a competent bacterial cell of the invention may be transformed by a vector of the invention using electroporation.
- Methods of making bacterial cells "competent" are well-known in the art, and typically involve preparation of the bacterial cells so that the cells take up exogenous DNA.
- methods of electroporation are known in the art, and detailed descriptions of such methods may be found, for example, in Sambrook et al. (1989, supra).
- the transformation of a competent cell with vector DNA may be also accomplished using chemical-based methods.
- One example of a well-known chemical-based method of bacterial transformation is described by Inoue, et al. (1990, Gene 96:23-28). Other methods of transformation will be known to the skilled artisan.
- a transformed host cell of the present invention may be used to express a truncated GalNAcT2 polypeptide of the present invention.
- a transformed host cell contains a vector of the invention, which contains therein a nucleic acid sequence encoding an truncated polypeptide of the invention.
- the truncated polypeptide is expressed using any expression method known in the art (for example, IPTG).
- IPTG IPTG
- the expressed truncated polypeptide may be contained within the host cell, or it may be secreted from the host cell into the growth medium.
- an expressed polypeptide that is secreted from a host cell may be isolated from the growth medium. Isolation of a polypeptide from a growth medium may include removal of bacterial cells and cellular debris. By way of another non-limiting example, an expressed polypeptide that is contained within a host cell may be isolated from the host cell. Isolation of such an "intracellular" expressed polypeptide may include disruption of the host cell and removal of cellular debris from the resultant mixture.
- Purification of a truncated polypeptide expressed in accordance with the present invention may be effected by any means known in the art. The skilled artisan will know how to determine the best method for the purification of a polypeptide expressed in accordance with the present invention. A purification method will be chosen by the skilled artisan based on factors such as, but not limited to, the expression host, the contents of the crude extract of the polypeptide, the size of the polypeptide, the properties of the polypeptide, the desired end product of the polypeptide purification process, and the subsequent use of the end product of the polypeptide purification process.
- isolation or purification of a truncated polypeptide expressed in accordance with the present invention may not be desired.
- an expressed polypeptide may be stored or transported inside the bacterial host cell in which the polypeptide was expressed.
- an expressed polypeptide may be used in a crude lysate form, which is produced by lysis of a host cell in which the polypeptide was expressed.
- an expressed polypeptide may be partially isolated or partially purified according to any of the methods set forth or described herein. The skilled artisan will know when it is not desirable to isolate or purify a polypeptide of the invention, and will be familiar with the techniques available for the use and preparation of such polypeptides.
- a eukaryotic host cell of the invention When armed with the disclosure set forth herein, the skilled artisan would also know how to prepare a eukaryotic host cell of the invention.
- an isolated nucleic acid encoding a truncated GalNAcT2 polypeptide may be introduced into a eukaryotic host cell, for example, using a lentivirus-based genomic integration or plasmid- based transfection (Sambrook et al., Third Edition, Molecular Cloning: A Laboratory Manual, Cold Spring Harbor Laboratory, New York (2001)).
- a eukaryotic host cell is a fungal cell.
- a nucleic acid encoding a truncated polypeptide of the invention is cloned into a lentiviral vector containing a specific promoter sequence for expression of the truncated polypeptide.
- the truncated polypeptide-containing lentiviral vector is then used to transfect a host cell for expression of the truncated polypeptide.
- a nucleic acid encoding a truncated polypeptide of the invention is introduced into a host cell using a viral expression system.
- Viral expression systems are well-known in the art, and will not be described in detail herein.
- a viral expression system is a mammalian viral expression system.
- a viral expression system is a baculo virus expression system. Such viral expression systems are typically commercially available from numerous vendors.
- Insect cells can also be used for expression of a truncated polypeptide of the present invention.
- Sf9, SF9 + , Sf21, High FiveTM or Drosophila Schneider S2 cells can be used.
- a baculovirus, or a baculovirus/insect cell expression system can be used to express a truncated polypeptide of the invention using a pAcGP67, pFastBac, pMelBac, or pIZ vector and a polyhedrin, plO, or O ⁇ IE3 actin promoter.
- a Drosophila expression system can be used with a pMT or pAC5 vector and an MT or Ac5 promoter.
- a truncated GalNAcT2 polypeptide of the invention can also be expressed in mammalian cells.
- 294, HeLa, HEK, NSO, Chinese hamster ovary (CHO), Jurkat, or COS cells can be used to express a truncated polypeptide of the invention.
- a suitable vector such as pT-Rex, pSecTag2, pBudCE4.1, or pCDNA/His Max vector can be used, along with, for example, a CMV promoter.
- mammalian cell culture systems can be employed to express recombinant protein.
- mammalian expression systems include the COS-7 lines of monkey kidney fibroblasts, described by Gluzman, Cell 23:175 (1981), and other cell lines capable of expressing a compatible vector, for example, the C127, 3T3, CHO, HeLa and BHK cell tines.
- Mammalian expression vectors may comprise an origin of replication, a suitable promoter and also any necessary ribosome binding sites, polyadenylation site, splice donor and acceptor sites, transcriptional termination sequences, and 5' flanking nontranscribed sequences.
- DNA sequences derived from the SV40 viral genome for example, SV40 origin, early promoter, enhancer, splice, and polyadenylation sites may be used to provide the required nontranscribed genetic elements.
- vector DNA can be introduced into a eukaryotic cell using conventional transfection techniques.
- transfection refers to a variety of art-recognized techniques for introducing foreign nucleic acid (e.g., DNA) into a host cell, including, DEAE-dextran-mediated transfection, lipofection, or electroporation.
- Suitable methods for transforming or transfecting host cells can be found in Sambrook, et al. (Molecular Cloning: A Laboratory Manual. 3nd ed., Cold Spring Harbor Laboratory, Cold Spring Harbor Laboratory Press, Cold Spring Harbor, N.Y., 2001), and other such laboratory manuals.
- a gene that encodes a selectable marker (e.g., resistance to antibiotics) is generally introduced into the host cells along with the gene of interest.
- selectable markers include those that confer resistance to drugs, such as G418, hygromycin and methotrexate.
- Nucleic acid encoding a selectable marker can be introduced into a host cell on the same vector as that encoding a truncated polypeptide of the invention or can be introduced on a separate vector. Cells stably transfected with the introduced nucleic acid can ' be identified by drug selection (e.g., cells that have incorporated the selectable marker gene will survive, while the other cells die).
- a truncated GalNAcT2 polypeptide of the present invention may be truncated in various ways, as would be known and understood by the skilled artisan, when armed with the present disclosure.
- Examples of truncated polypeptides of the present invention include, but are not limited to, a polypeptide lacking a single N-terminal residue, a polypeptide lacking a single C-terminal residue, a polypeptide lacking both an single N-terminal residue and a single C-terminal residue, a polypeptide lacking a contiguous sequence of residues from the N-terminus, a polypeptide lacking a contiguous sequence of residues from the C-terminus, and any such combinations thereof.
- a full-length human GalNAcT2 polypeptide may contain one or more identifyable polypeptide domains in addition to the "active domain," the domain primarily responsible for the catalytic activity, of GalNAcT2. This is because it is known in that art that a full-length GalNAcT2 polypeptide, and in particular, a full-length human GalNAcT2 polypeptide, contains a signal domain, a transmembrane domain, and a stem domain, in addition to an active domain.
- a full-length human GalNAcT2 may have a signal domain at the amino-terminus of the polypeptide, followed by a transmembrane domain immediately adjacent to the signal domain, followed by a stem domain that is immediately adjacent to the transmembrane domain, followed by an active domain that extends to the carboxy-terminus of the polypeptide and is located immediately adjacent to the stem domain.
- a GalNAcT2 polypeptide of the invention is a truncated human GalNAcT2 polypeptide lacking all or a portion of the GalNAcT2 signal domain.
- a GalNAcT2 polypeptide of the invention is a truncated human GalNAcT2 polypeptide lacking the GalNAcT2 signal domain and all or a portion of the GalNAcT2 transmembrane domain.
- a GalNAcT2 polypeptide of the invention is a truncated human GalNAcT2 polypeptide lacking the GalNAcT2 signal domain, the GalNAcT2 transmembrane domain and all or a portion the GalNAcT2 stem domain.
- a truncated GalNAcT2 mutant of the present invention is based on the point at which the full-length polypeptide is truncated.
- a " ⁇ 40 human truncated GalNAcT2" mutant of the invention refers to a truncated GalNAcT2 polypeptide of the invention in which amino acids 1 through 40, counting from the N-terminus of the full-length polypeptide, are deleted from the polypeptide. Therefore, the N-terminus of the ⁇ 40 human truncated GalNAcT2 mutant begins with the amino acid residue that would be referred to as "amino acid 41" of the full- length polypeptide.
- amino acid 41 amino acid residue that would be referred to as "amino acid 41" of the full- length polypeptide.
- the present invention therefore also includes an isolated polypeptide comprising a truncated GalNAcT2 polypeptide.
- an isolated truncated GalNAcT2 polypeptide of the present invention has at least about 90% identity to a polypeptide having the amino acid sequence of any one of the sequences set forth in SEQ ID NO:4, SEQ ID NO: 8 or SEQ JJD NO: 10.
- the isolated polypeptide is about 95% identical, and even more preferably, about 98% identical, still more preferably, about 99% identical, and most preferably, the isolated polypeptide comprising a truncated GalNAcT2 polypeptide is identical to the polypeptide set forth in one of SEQ ID NO:4, SEQ ID NO: 8 or SEQ ID NO:10.
- the present invention also provides for analogs of polypeptides which comprise a truncated GalNAcT2 polypeptide as disclosed herein. Analogs can differ from naturally occurring proteins or peptides by conservative amino acid sequence differences or by modifications which do not affect sequence, or by both.
- conservative amino acid changes may be made, which although they alter the primary sequence of the protein or peptide, do not normally alter its function.
- Conservative amino acid substitutions typically include substitutions within the following groups: glycine, alanine; valine, isoleucine, leucine; aspartic acid, glutamic acid; asparagine, glutamine; serine, threonine; lysine, arginine; phenylalanine, tyrosine.
- Modifications include in vivo, or in vitro chemical derivatization of polypeptides, e.g., acetylation, or carboxylation. Also included are modifications of glycosylation, e.g., those made by modifying the glycosylation patterns of a polypeptide during its synthesis and processing or in further processing steps; e.g., by exposing the polypeptide to enzymes which affect glycosylation, e.g., mammalian glycosylating or deglycosylating enzymes. Also embraced are sequences which have phosphorylated amino acid residues, e.g., phosphotyrosine, phosphoserine, or phosphothreonine.
- polypeptides which have been modified using ordinary molecular biological techniques so as to improve their resistance to proteolytic degradation or to optimize solubility properties or to render them more suitable as a therapeutic agent.
- Analogs of such polypeptides include those containing residues other than naturally occurring L- amino acids, e.g., D-amino acids or non-naturally occurring synthetic amino acids.
- the peptides of the invention are not limited to products of any of the specific exemplary processes listed herein.
- Fragments of a truncated GalNAcT2 polypeptide of the invention are included in the present invention, provided the fragment possesses the biological activity of the full- length polypeptide. That is, a truncated GalNAcT2 polypeptide of the present invention can catalyze the same glycosyltransfer reaction as the full-length GalNAcT2.
- a truncated human GalNAcT2 polypeptide has the ability to transfer a GalNAc moiety from a UDP-GalNAc donor to a granulocyte-colony stimulating factor (G- CSF) acceptor, wherein such a transfer results in the O-linked covalent coupling of a GalNAc moiety to a threonine residue of G-CSF. Therefore, a smaller than full-length, or "truncated,” GalNAcT2 is included in the present invention provided that the truncated GalNAcT2 has GalNAcT2 biological activity.
- G- CSF granulocyte-colony stimulating factor
- compositions comprising an isolated truncated GalNAcT2 polypeptide as described herein may include highly purified truncated GalNAcT2 polypeptides.
- compositions comprising truncated GalNAcT2 polypeptides may include cell lysates prepared from the cells used to express the particular truncated GalNAcT2 polypeptides.
- truncated GalNAcT2 polypeptides of the present invention may be expressed in one of any number of cells suitable for expression of polypeptides, such cells being well-known to one of skill in the art, as described in detail elsewhere herein.
- Substantially pure protein isolated and obtained as described herein may be purified by following known procedures for protein purification, wherein an immunological, enzymatic or other assay is used to monitor purification at each stage in the procedure.
- Protein purification methods are well known in the art, and are described, for example in Deutscher et al. (ed., 1990, Guide to Protein Purification, Harcourt Brace Jovanovich, San Diego).
- the present invention features a method of expressing a truncated polypeptide.
- Polypeptides which can be expressed according to the methods of the present invention include a truncated GalNAcT2 polypeptide. More preferably, polypeptides which can be expressed according to the methods of the present invention include, but are not limited to, a truncated human GalNAcT2 polypeptide. In a preferred embodiment, a polypeptide which can be expressed according to the methods of the present invention is a polypeptide comprising any one of the polypeptide sequences set forth in SEQ ID NO:4, SEQ ID NO:8 or SEQ ID NO: 10.
- the present invention features a method of expressing a truncated GalNAcT2 polypeptide encoded by an isolated nucleic acid of the invention, as described elsewhere herein, wherein the expressed truncated GalNAcT2 polypeptide has the property of catalyzing the transfer of a GalNAc moiety to an acceptor moiety.
- a method of expressing a truncated GalNAcT2 polypeptide includes the steps of cloning an isolated nucleic acid of the invention into an expression vector, inserting the expression vector construct into a host cell, and expressing a truncated GalNAcT2 polypeptide therefrom.
- Methods of expression of polypeptides are discussed in extensive detail elsewhere herein. Methods of expression of a truncated polypeptide of the present invention will be understood to include, but not to be limited to, all such methods as described herein.
- the truncated GalNAcT2 polypeptides of the invention are expressed as insoluble proteins, e.g., in an inclusion protein in a bacterial host cell. Methods of refolding insoluble glycosyltransferases, including GalNAcT2 polypeptides, are disclosed in U.S. Provisional Patent Application Serial No. 60/542,210, filed February 4, 2004; U.S.
- the present invention also features a method of catalyzing the transfer of a GalNAc moiety to a GalNAc acceptor moiety, wherein the GalNAc-transfer reaction is carried out by incubating a truncated GalNAcT2 polypeptide of the invention with a GalNAc donor moiety and a GalNAc acceptor moiety.
- a truncated GalNAcT2 polypeptide of the invention mediates the covalent linkage of a GalNAc moiety to a GalNAc acceptor moiety, thereby catalyzing the transfer of a GalNAc moiety to an acceptor moiety.
- a truncated GalNAcT2 polypeptide useful in a glycosyltransfer reaction is a truncated human GalNAcT2 polypeptide.
- the human GalNAc T2 glycosyltransfer reaction involves the transfer of a GalNAc residue from a GalNAc donor to a GalNAc acceptor.
- a method of catalyzing the transfer of a GalNAc moiety to an acceptor moiety includes the steps of incubating a truncated GalNAcT2 polypeptide with UDP-GalNAc GalNAc donor and a granulocyte colony stimulating factor (G-CSF) acceptor moiety, wherein the truncated GalNAcT2 polypeptide mediates the transfer of GalNAc from the UDP-GalNAc donor to the GCSF acceptor.
- G-CSF granulocyte colony stimulating factor
- the present invention also features a polypeptide acceptor moiety.
- a polypeptide acceptor moiety is a human growth hormone.
- a polypeptide acceptor moiety is an erythropoietin.
- a polypeptide acceptor moiety is an interferon- alpha.
- a polypeptide acceptor moiety is an interferon-beta.
- a polypeptide acceptor moiety is an interferon-gamma.
- a polypeptide acceptor moiety is a lysosomal hydrolase.
- a polypeptide acceptor moiety is a blood factor polypeptide.
- a polypeptide acceptor moiety is an anti-tumor necrosis factor-alpha.
- a polypeptide acceptor moiety is follicle stimulating hormone.
- the present invention also features a method of transferring a GalNAc-polyethyleneglycol conjugate to an acceptor molecule
- an acceptor molecule is a polypeptide.
- an acceptor molecule is a glycopeptide.
- Compositions and methods useful for designing, producing and transferring a GalNAc- polyethyleneglycol conjugate to an acceptor molecule are discussed at length in International (PCT) Patent Application No. WO03/031464 (PCT/US02/32263) and U.S. Patent Application No. 2004/0063911, each of which is incorporated herein by reference in its entirety. Methods of assaying for glycosyltransferase activity are well-known in the art.
- Example 1 Cloning, Expression, and Refolding of Human Polypeptide N- acetylgalactosaminyltransferase II (GalNAcT2 in E. coli JM109
- Truncated human polypeptide N-acetylgalactosaminyltransferase II was expressed as maltose binding protein (MBP)-fusion proteins in inclusion bodies from E. coli JM109 cells.
- MBP maltose binding protein
- the production of active enzyme was examined by refolding and assaying against two polypeptide acceptors. Therefore, described herein is the generation of several truncated forms of human polypeptide GalNAcT2 as maltose binding protein fusion proteins in E.coli JM109 cells.
- the recombinant proteins are refolded from isolated inclusion bodies using the Hampton Foldlt screen kit (Hampton Research, Aliso Vieja, CA). All four constructs were expressed in JM109 E.coli at levels of approximately 2g/L culture media.
- PCR Polymerase Chain Reaction
- amplifications were performed in a final reaction volume of 50 ⁇ l containing 5 ⁇ l of template DNA (11 ⁇ g/ml, 100-fold diluted pBKS-Full ppGalNAcT2), 40 pmol of 5'- primer and 3'- primer, 10 nmol of dNTP mixture, and 5 units of HerculaseTM Enhanced DNA Polymerase under the conditions of 31 cycles of denaturation at 95°C for 45 seconds, annealing at 62°C for 45 seconds, and extension at 74°C for 170 seconds.
- PCR products were subjected to 1% agarose gel elecfrophoresis. DNA fragments were excised and purified by QIAEX II gel extraction kit (Qiagen, Valencia, CA). Table 1 illustrates the primers used in the PCR reactions.
- JM109 cells were cultured in a 15 ml culture tube containing 6 ml LB medium and 15 ⁇ g/ml of kanamycin overnight at 37°C with rapid shaking (250 rpm). For each culture, two milliliters of starting culture was transferred to a 50 ml centrifuge tube containing 23 ml LB medium with 15 ⁇ g/ml kanamycin and incubated at 37°C with rapid shaking for 3 hours. Isopropyl-1-thio- ⁇ -D-galactopyranoside (IPTG) was added to a final concentration of 0.4 mM to induce the protein expression.
- IPTG Isopropyl-1-thio- ⁇ -D-galactopyranoside
- Each sample for SDS-PAGE separation was prepared by mixing 5 ⁇ l of whole cells suspension, lysate, or inclusion bodies suspension with 5 ⁇ l of 2 x Tris-Glycine SDS sample buffer and 1.1 ⁇ l of DTT (1 M). The mixture was heated at 98°C for 5 minutes, cooled to room temperature, and loaded to each well of a 1.0 mm x 15 well 4-20% Tris-Glycine gradient gel. The elecfrophoresis was conducted at 120 V for 100 minutes. The gel was then stained for 2 hours and de-stained with distilled water (see Figures 4-6).
- Inclusion bodies were dissolved at 20 mg/ml (high protein concentration) or 2 mg/ml concentration (low protein concentration) in solubilizing buffer containing 4 M Guanidine-HCl, 100 mM Tris-HCl, pH 9.0, 5 mM EDTA, and 10 mM DTT. Refolding of inclusion bodies by Hampton Foldlt Screen Kit was carried out by following the manufacturer's protocol, except that a 10-fold less volume was used (100 ⁇ l -scale) (Hampton Products, Aliso Viejo, CA).
- Non-radioactive enzyme activity assays for lysates were carried out in a 0.5 ⁇ l microcentrifuge tube at 37°C for overnight in a final volume of 10 ⁇ l containing 50 mM MES buffer, pH 6.0, MnCl 2 (15 mM), MgCl 2 (15 mM), NaCl (0.15 M), UDP-GalNAc (5 mM), 1.5 ⁇ g G-CSF (acceptor), and 2.15 ⁇ l of lysate sample. Enzyme was substituted by H 2 O as a negative control. Purified recombinant ppGalNAcT2 (0.5 ⁇ l) from Sf9 baculovirus expression system was used as the positive control.
- DNA fragments for ppGalNAcT2 genes were successfully amplified by PCR as shown in Figure 5.
- Vector plasmid DNA pCWin2MBP was digested by BamHI and Xhol, and purified on a 1% agarose gel. The gel purified DNA fragment was digested by the same two enzymes and purified. After digestion, the DNA fragments were clean as visualized on an agarose gel ( Figure 2B).
- BamHI and Xhol digestion of the plasmids purified from the selected twelve colonies showed predicted correct pattern on a 1% agarose gel. The size of the vector was around 6.2 kb, and the inserts were approximately 1.5 kb.
- Maltose-binding protein (MBP) expressed in the JM109 transformed with ⁇ CWin2MBP vector plasmid showed a band at around 43 kDa. Over 90% of the proteins in the whole cells are MBP. The #2 colony of the construct N41R expressed a shorter protein than expected, indicating the occurrence of mutation. All other eleven colonies showed a band at about 100 kDa for MBP-ppGalNAcT2 fusion proteins, with over 80% of the total proteins were the target fusion proteins.
- Refolding experiments on MBP-GalNAcT2 were carried out on a 1 ml scale, with four different MBP-GalNAcT2 DNA constructs and under 16 different possible refolding conditions. Refolding was performed using the Hampton Research Foldit kit (Hampton Research, Aliso Viejo, CA) and the assays were performed via radioactive detection of [ 3 H] UDP-GalNAc addition to a MuC-2 peptide and via matrix-assisted laser desorption ionization mass spectrometry (MALDI) analysis utilizing addition of GalNAc to Interferon ⁇ -2b and G- CSF.
- MALDI matrix-assisted laser desorption ionization mass spectrometry
- GalNAcT2 constructs used in the present invention comprised DNA encoding various amino terminal amino acid truncation mutants of the original human GalNAcT2 protein, including the following constructs, which begin with the N-terminal amino acid as indicated:
- Martone L-Broth containing lO ⁇ g/ml Kanamycin sulfate with a pipette tip scraping from the particular glycerol stock culture. This procedure was performed on all four constructs for a total of four starter cultures. Starter cultures were incubated overnight at 37°C, with rotary shaking at 250rpm. From the overnight cultures, four 275 ml Martone L-Broth cultures containing lO ⁇ g/ml Kanamycin sulfate were prepared. Each of these cultures was inoculated with 275 ⁇ L of one of the 2 ml starter cultures of constructs 1 through 4. These 275 ml cultures were incubated overnight at 37°C, with shaking at 250rpm.
- IL Martone L-Broth cultures containing 1 O ⁇ g/ml Kanamycin sulfate were prepared. Each of these cultures was inoculated with 40 ml of one of the 275 ml cultures of constructs 1 though 4. These IL cultures were incubated at 37°C, with shaking at 250rpm, until the OD600 measured approximately 1.0. Upon reaching this point, IPTG was added to each of the four IL cultures to a final concentration of 0.4mM. Cultures were then allowed incubate overnight at 37°C, with shaking at 250rpm.
- MBP-GalNAcT2 refold samples were purified by use of G-50 Macro Spin Columns (Harvard Bioscience, Holliston, MA). Caps were removed from the G-50 columns and columns were placed into 2 ml microcentrifuge tubes.
- H 2 O 500 ⁇ l was added to each column and the columns were allowed to incubate for 15 minutes to hydrate. The columns were then centrifuged at ⁇ 2000 x g for 4 minutes after which they were transferred to new 2 ml centrifuge tubes. Each refold solution (150 ⁇ l) was applied to one of the columns. Columns were then centrifuged at 2000 x g for ⁇ 2 minutes. Resulting permeates represented the purified refold samples.
- a radiolabeled [ H]-UDP-GalNAc assay was performed to determine the activity of the E.coli-expressed refolded MBP-GalNAcT2 by monitoring the addition of radiolabeled GalNAc to a peptide acceptor.
- the acceptor was a MuC-2 - like peptide having the sequence MVTPTPTPTC (SEQ ID NO: 16).
- the initial screen was performed on refolded protein samples which had been purified by dialysis. Subsequent refold samples were freshly refolded and purified by G-50 gel filtration.
- the assay included protein refold samples, GalNAcT2 from Baculovirus as a positive control, a negative control sample with all the components except enzyme and a maximum input sample which contained all components except enzyme. A total of 19 samples were tested.
- the assay solution consisted of the components listed in Table 3:
- the assay solution was prepared as shown in Table 4 for each reaction.
- the above assay was performed to determine whether E.coli-expressed refolded MBP-GalNAcT2 could transfer GalNAc to G-CSF acceptor from a UDP-GalNAc donor.
- construct 2 in refold buffer 8 was assayed for GalNAcT2 activity.
- GalNAcT2 from Baculovirus was assayed.
- the assay solution was prepared for each reaction as shown in Table 5.
- Table 5 Parameters for G-CSF acceptor GalNAcT2 activity assay
- MBP-GalNAcT2 The expression of MBP-GalNAcT2 was observed by way of the SDS-Page gel analysis of JM109 pCWin2 MBP-GalNAcT2 whole cell samples before and after induction by IPTG (Figure 7).
- the protein gel shows a clear increase in protein expression in the induced state compared to the uninduced state. Furthermore there is a distinct band at -lOOkDa that substantially increases after induction which correlates to the expected size of the MBP-GalNAcT2 band.
- Protein samples were diluted by combining 950 ⁇ L of H 2 O with 50 ⁇ L of protein sample. Samples were then analyzed using a UV spectrophotometer. Protein concentration was calculated from absorption values and the molar extinction coefficients: Construct 1 - 0.65mg/ml per 1 A 28 o unit, Construct 2 - 0.64mg/ml per 1 A 280 unit, as shown in Table 7.
- construct 2 was tested under refold conditions 3, 8, 11, 12, 15 and 16 from the Hampton Foldit kit (Hampton Research, Aliso Viejo, CA). These refolded enzymes were purified by G-50 gel filtration and then tested for activity by the radioactive assay. Results indicate that after overnight incubation on a rotator, greatest activity was obtained from refold condition 15.
- construct 2 was tested under refold conditions 3, 8, 11, 12, 15 and 16 from the Hampton Foldit kit (Hampton Research, Aliso Viejo, CA) after being rotated overnight at 4°C and left resting at 4°C for 5 days. These refolded enzymes were purified by G-50 gel filtration and then tested for activity by the radioactive assay. Results indicated that after 5 days in refold buffer 8, construct 8 displayed the highest activity. Therefore it was determined that conditions 8 and 15 had the greatest potential for producing a properly folded and active MBP-GalNAcT2.
- An IF ⁇ -2b assay was performed on overnight refolds of constructs 1 and 2 in refold buffer 15 (1-15 and 2-15, respectively) and was incubated at 32°C for 5 days. Time points were taken of the IF ⁇ -2b reaction at 16 hours and 5 days. The results indicate that the parental peak for IF ⁇ -2b is at MW -19267. A successful reaction would be indicated by addition of -203 molecular weight to that peak. From the 5 day data for refolds 1-15 and 2- 15, a developing peak was observed at -119478 and -19473 respectively, a difference of approximately 203 MW.
- a G-CSF assay was performed on the 5-day refolded enzymes of construct 2 in refold buffer 8. The G-CSF reaction was allowed to incubate at 32°C for 4 days. The reaction was analyzed at the 4 day time point. The parental peak for G-CSF is expected at MW -18786. A successful reaction would be indicated by addition of -203 molecular weight to that peak. From the 3 day data for refolded enzymes 2-8, a developing peak was observed at -19001, a difference of approximately 203 MW. This data again indicated that GalNAc was added to G-CSF by the refolded GalNAcT2 protein and confirmed what was reported by the radioactive assay and the IF ⁇ -2b assay as reported elsewhere herein.
- GalNAcT2 truncation mutants of the present invention are also useful for the transfer of a glycosyl-polyethyleneglycol ("glycosyl-PEG") conjugate to a polypeptide, also known as "glycoPEGylation" of a polypeptide.
- glycosyl-PEG glycosyl-polyethyleneglycol
- glycoPEGylation also known as "glycoPEGylation" of a polypeptide.
- SA GalNAc-sialic acid
- a glycoPEGylation reaction mixture was prepared in order to glycoPEGylate G- CSF.
- the reaction mixture contained 5 ⁇ l of ⁇ 51 GalNAcT2-MBP (20 ⁇ U), 2 ⁇ l of GalNAc- ⁇ 2,6-sialyltransferase (ST6GalNAcI), 6.25 mM MnCl 2 , 15 mM UDP-GalNAc, 0.75 mM CMP-SA-PEG (20K), and between 2 ⁇ l and 10 ⁇ l of 2 mg/ml G-CSF.
- Gel elecfrophoresis of the reaction products demonstrated that ⁇ 51 GalNAcT2-MBP transferred a GalNAc-sialic acid (SA)-PEG conjugate to G-CSF (Figure 9).
- Example 3 Optimization of Purification and Refolding of ⁇ 51 GalNAcT2-MBP
- ⁇ 51 GalNAcT2 refolding and purification development as set forth herein demonstrates the utility of a two column purification procedure for purification of GalNAcT2 mutants.
- Q Sepharose Fast Flow in binding mode and Q Sepharose XL in binding and flow through mode as an initial purification step has been explored.
- Q Sepharose XL in flow through mode using a NaCl concentration of lOOmM in the load led to best recovery and purity of active ⁇ 51 GalNAcT2-MBP.
- the use of Hydroxyapatite Type I has been considered as a second column step.
- Initial data indicate ⁇ 51 GalNAcT2-MBP binds to this resin and can be eluted as an active enzyme with a phosphate gradient.
- ⁇ 51 GalNAcT2-MBP was cloned and expressed as set forth elsewhere herein.
- DWIBs double- washed inclusion bodies
- harvested cell pellet was resuspended in lOmM Tris/ 5mM EDTA pFI 7.5 (5mL/g cells) and lysed in two passes using a micro fluidizer at 12,000psi.
- Inclusion bodies were harvested by centrifugation at 6,000 rpm for 20 min in a Sorvall RC-3B. The pellet was washed twice by resuspension in above buffer at 5mL/g pellet followed by centrifugation at 6,000 RPM for 20min. DWIBs were aliquoted and stored at -20°C.
- ⁇ 51 GalNAcT2-MBP refolds were performed by solubilizing 2.5g of DWIB's in 250 mL of 7M urea/ 50mM Tris/ lOmM DTT/ 5mM EDTA pH 8.0 at 4°C. 50mL solubilized ⁇ 51 GalNAcT2-MBP DWIB's were added to IL of refold buffer at 4°C while stirring (21- fold dilution - 0.5mg/mL). Refolding was allowed to proceed for 20.5h at 4°C with stirring.
- Refolds were filtered using a Cuno Zeta Plus BioCap (Cuno, Meriden, CT), concentrated 4-fold and diafiltered on a 1 ft2 30kDa MWCO TFF (regenerated cellulose) filter at constant volume with 5 diavolumes of lOmM Tris/ 5mM NaCl pH 8.
- ⁇ 51 GalNAcT2-MBP bound tightly to QSFF resin under above conditions with 5mM NaCl in load and equilibration buffers. Active ⁇ 51 GalNAcT2-MBP eluted at the beginning of the major peak and appears as a doublet on a nonreduced 4-20% Tris-glycine gel.
- the major contaminant is a currently unidentified band running at a slightly lower molecular weight close to the 98kDa marker band. A variety of other contaminants elute with inactive ⁇ 51 GalNAcT2-MBP in the remainder of the major peak.
- ⁇ 51 GalNAcT2-MBP bound tightly to QXL resin if the same conditions as for QSFF binding were applied (i.e. 5mM NaCl). Increasing ⁇ 51 GalNAcT2-MBP activity was observed in flow through and wash at higher NaCl concentrations in the load. Interestingly, the major contaminating band observed in QSFF purification was not visible in the flow through if the load contained 50 and lOOmM NaCl. At both NaCl concentrations the majority of active ⁇ 51 GalNAcT2-MBP could be found in flow through and wash; only some residual ⁇ 51 GalNAcT2-MBP activity was detected in the left shoulder of the elution peak.
- Hydroxyapatite Type I 80 ⁇ m (BioRad, Hercules, CA) was examined as a second column step. Active ⁇ 51 GalNAcT2-MBP partially purified over QSFF (using bind and elute mode) was used to investigate if active ⁇ 51 GalNAcT2-MBP would bind to an HA Type I resin and would be useful to further purify the protein. For this purpose, a 2.25 mL HA Type I column was pre-equilibrated with 5mM NaPO4/ 5mM NaCl pH 7.0 (C). Active ⁇ 51 GalNAcT2-MBP eluted from QSFF was adjusted to pH 7.0 with IM HC1 and applied onto the HA Type I column.
- the protein was eluted using a 20 CV gradient from 0-50% 300mM NaPO4/ 5mM NaCl pH 7.0 (D), followed by a 5 CV gradient from 50-100%) D.
- the column was regenerated using 0.5M NaOH.
- the data obtained indicate that ⁇ 51 GalNAcT2-MBP binds to hydroxyapatite type I resin and can be eluted as an active enzyme.
Landscapes
- Health & Medical Sciences (AREA)
- Life Sciences & Earth Sciences (AREA)
- Chemical & Material Sciences (AREA)
- Organic Chemistry (AREA)
- Engineering & Computer Science (AREA)
- Bioinformatics & Cheminformatics (AREA)
- Wood Science & Technology (AREA)
- Medicinal Chemistry (AREA)
- General Health & Medical Sciences (AREA)
- Genetics & Genomics (AREA)
- Zoology (AREA)
- General Engineering & Computer Science (AREA)
- Pharmacology & Pharmacy (AREA)
- Biochemistry (AREA)
- Biotechnology (AREA)
- Biomedical Technology (AREA)
- Molecular Biology (AREA)
- Chemical Kinetics & Catalysis (AREA)
- General Chemical & Material Sciences (AREA)
- Nuclear Medicine, Radiotherapy & Molecular Imaging (AREA)
- Microbiology (AREA)
- Animal Behavior & Ethology (AREA)
- Public Health (AREA)
- Veterinary Medicine (AREA)
- Micro-Organisms Or Cultivation Processes Thereof (AREA)
- Preparation Of Compounds By Using Micro-Organisms (AREA)
- Peptides Or Proteins (AREA)
- Medicines That Contain Protein Lipid Enzymes And Other Medicines (AREA)
- Enzymes And Modification Thereof (AREA)
Applications Claiming Priority (3)
| Application Number | Priority Date | Filing Date | Title |
|---|---|---|---|
| US57653004P | 2004-06-03 | 2004-06-03 | |
| US59858404P | 2004-08-03 | 2004-08-03 | |
| PCT/US2005/019442 WO2005121331A2 (en) | 2004-06-03 | 2005-06-03 | Truncated galnact2 polypeptides and nucleic acids |
Publications (2)
| Publication Number | Publication Date |
|---|---|
| EP1765992A2 true EP1765992A2 (de) | 2007-03-28 |
| EP1765992A4 EP1765992A4 (de) | 2009-01-07 |
Family
ID=35945309
Family Applications (1)
| Application Number | Title | Priority Date | Filing Date |
|---|---|---|---|
| EP05758682A Withdrawn EP1765992A4 (de) | 2004-06-03 | 2005-06-03 | Verkürzte galnact2-polypeptide und nukleinsäuren |
Country Status (3)
| Country | Link |
|---|---|
| EP (1) | EP1765992A4 (de) |
| JP (1) | JP2008512085A (de) |
| WO (1) | WO2005121331A2 (de) |
Families Citing this family (35)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| US7214660B2 (en) | 2001-10-10 | 2007-05-08 | Neose Technologies, Inc. | Erythropoietin: remodeling and glycoconjugation of erythropoietin |
| US7795210B2 (en) | 2001-10-10 | 2010-09-14 | Novo Nordisk A/S | Protein remodeling methods and proteins/peptides produced by the methods |
| US7173003B2 (en) | 2001-10-10 | 2007-02-06 | Neose Technologies, Inc. | Granulocyte colony stimulating factor: remodeling and glycoconjugation of G-CSF |
| US7696163B2 (en) | 2001-10-10 | 2010-04-13 | Novo Nordisk A/S | Erythropoietin: remodeling and glycoconjugation of erythropoietin |
| US7157277B2 (en) | 2001-11-28 | 2007-01-02 | Neose Technologies, Inc. | Factor VIII remodeling and glycoconjugation of Factor VIII |
| JP2006523211A (ja) | 2003-03-14 | 2006-10-12 | ネオス テクノロジーズ インコーポレイテッド | 分岐水溶性ポリマーとその複合物 |
| WO2006127896A2 (en) | 2005-05-25 | 2006-11-30 | Neose Technologies, Inc. | Glycopegylated factor ix |
| US8791070B2 (en) | 2003-04-09 | 2014-07-29 | Novo Nordisk A/S | Glycopegylated factor IX |
| NZ542586A (en) | 2003-04-09 | 2009-11-27 | Biogenerix Ag | Method of forming a covalent conjugate between a G-CSF peptide and PEG |
| AU2004240553A1 (en) | 2003-05-09 | 2004-12-02 | Neose Technologies, Inc. | Compositions and methods for the preparation of human growth hormone glycosylation mutants |
| US9005625B2 (en) | 2003-07-25 | 2015-04-14 | Novo Nordisk A/S | Antibody toxin conjugates |
| US8633157B2 (en) | 2003-11-24 | 2014-01-21 | Novo Nordisk A/S | Glycopegylated erythropoietin |
| US20080305992A1 (en) | 2003-11-24 | 2008-12-11 | Neose Technologies, Inc. | Glycopegylated erythropoietin |
| US20060040856A1 (en) | 2003-12-03 | 2006-02-23 | Neose Technologies, Inc. | Glycopegylated factor IX |
| US7956032B2 (en) | 2003-12-03 | 2011-06-07 | Novo Nordisk A/S | Glycopegylated granulocyte colony stimulating factor |
| US7338933B2 (en) | 2004-01-08 | 2008-03-04 | Neose Technologies, Inc. | O-linked glycosylation of peptides |
| WO2006010143A2 (en) | 2004-07-13 | 2006-01-26 | Neose Technologies, Inc. | Branched peg remodeling and glycosylation of glucagon-like peptide-1 [glp-1] |
| EP1799249A2 (de) | 2004-09-10 | 2007-06-27 | Neose Technologies, Inc. | Glycopegyliertes interferon alpha |
| EP3061461A1 (de) | 2004-10-29 | 2016-08-31 | ratiopharm GmbH | Remodellierung und glykopegylierung von fibroblasten-wachstumsfaktor (fgf) |
| US9029331B2 (en) | 2005-01-10 | 2015-05-12 | Novo Nordisk A/S | Glycopegylated granulocyte colony stimulating factor |
| EP1871795A4 (de) | 2005-04-08 | 2010-03-31 | Biogenerix Ag | Zusammensetzungen und verfahren zur herstellung von glycosylierungsmutanten eines proteaseresistenten menschlichen wachstumshormons |
| WO2006127910A2 (en) | 2005-05-25 | 2006-11-30 | Neose Technologies, Inc. | Glycopegylated erythropoietin formulations |
| US20070105755A1 (en) | 2005-10-26 | 2007-05-10 | Neose Technologies, Inc. | One pot desialylation and glycopegylation of therapeutic peptides |
| WO2007056191A2 (en) | 2005-11-03 | 2007-05-18 | Neose Technologies, Inc. | Nucleotide sugar purification using membranes |
| US20080242607A1 (en) | 2006-07-21 | 2008-10-02 | Neose Technologies, Inc. | Glycosylation of peptides via o-linked glycosylation sequences |
| US8969532B2 (en) | 2006-10-03 | 2015-03-03 | Novo Nordisk A/S | Methods for the purification of polypeptide conjugates comprising polyalkylene oxide using hydrophobic interaction chromatography |
| CN101796063B (zh) | 2007-04-03 | 2017-03-22 | 拉蒂奥法姆有限责任公司 | 使用糖聚乙二醇化g‑csf的治疗方法 |
| WO2008143944A2 (en) * | 2007-05-14 | 2008-11-27 | Government Of The Usa, As Represented By The Secretary, Department Of Health And Human Services | Methods of glycosylation and bioconjugation |
| US9493499B2 (en) | 2007-06-12 | 2016-11-15 | Novo Nordisk A/S | Process for the production of purified cytidinemonophosphate-sialic acid-polyalkylene oxide (CMP-SA-PEG) as modified nucleotide sugars via anion exchange chromatography |
| US8207112B2 (en) | 2007-08-29 | 2012-06-26 | Biogenerix Ag | Liquid formulation of G-CSF conjugate |
| JP5358840B2 (ja) * | 2007-11-24 | 2013-12-04 | 独立行政法人産業技術総合研究所 | Gfp(緑色蛍光蛋白質)の機能賦活・回復方法 |
| DK2257311T3 (da) | 2008-02-27 | 2014-06-30 | Novo Nordisk As | Konjugerede Faktor VIII-Molekyler |
| CN103163300B (zh) * | 2013-03-01 | 2015-02-11 | 苏州大学 | 人n-乙酰氨基半乳糖转移酶2双抗体夹心elisa试剂盒 |
| WO2018211529A1 (en) * | 2017-05-19 | 2018-11-22 | Council Of Scientific & Industrial Research | A method for producing refolded recombinant humanized ranibizumab |
| CN114058602B (zh) * | 2020-07-30 | 2023-08-22 | 中国中医科学院中药研究所 | 新疆紫草咖啡酸及迷迭香酸糖基转移酶及编码基因与应用 |
Family Cites Families (3)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| US7892730B2 (en) * | 2000-12-22 | 2011-02-22 | Sagres Discovery, Inc. | Compositions and methods for cancer |
| JP2005530309A (ja) * | 2002-03-27 | 2005-10-06 | モレックス インコーポレーテッド | 改善された保持性能を有する差動信号コネクタ組立体 |
| EP1558728B1 (de) * | 2002-11-08 | 2007-06-27 | Glycozym ApS | Verfahren zur identifizierung von die funktionen von polypeptid-galnac-transferasen modulierenden agentien, solche agentien umfassende pharmazeutische zusammensetzungen und verwendung solcher agentien zur herstellung von arzneimitteln |
-
2005
- 2005-06-03 EP EP05758682A patent/EP1765992A4/de not_active Withdrawn
- 2005-06-03 JP JP2007515582A patent/JP2008512085A/ja active Pending
- 2005-06-03 WO PCT/US2005/019442 patent/WO2005121331A2/en not_active Ceased
Non-Patent Citations (8)
| Title |
|---|
| BENNETT E P ET AL: "cDNA cloning and expression of a novel human UDP-N-acetyl-alpha-D-galactosamine. Polypeptide N-acetylgalactosaminyltransferase, GalNAc-t3" JOURNAL OF BIOLOGICAL CHEMISTRY, AMERICAN SOCIETY OF BIOLOCHEMICAL BIOLOGISTS, BIRMINGHAM,; US, vol. 271, no. 29, 19 July 1996 (1996-07-19), pages 17006-17012, XP002386961 ISSN: 0021-9258 * |
| BENNETT ERIC PAUL ET AL: "Cloning of a human UDP-N-acetyl-alpha-D-galactosamine:polypep tide N-acetylgalactosaminlytransferase that complements other GalNAc-transferases in complete O-glycosylation of the MUC1 tandem repeat" JOURNAL OF BIOLOGICAL CHEMISTRY, vol. 273, no. 46, 13 November 1998 (1998-11-13), pages 30472-30481, XP002504533 ISSN: 0021-9258 * |
| BOEGGEMAN E E ET AL: "The N-terminal stem region of bovine and human beta1,4-galactosyltransferase I increases the in vitro folding efficiency of their catalytic domain from inclusion bodies" PROTEIN EXPRESSION AND PURIFICATION, ACADEMIC PRESS, SAN DIEGO, CA, vol. 30, no. 2, 1 August 2003 (2003-08-01), pages 219-229, XP004439531 ISSN: 1046-5928 * |
| DONADIO S ET AL: "Recognition of cell surface acceptors by two human alpha-2,6-sialyltransferases produced in CHO cells" BIOCHIMIE, MASSON, PARIS, FR, vol. 85, no. 3-4, 1 January 2003 (2003-01-01), pages 311-321, XP002397070 ISSN: 0300-9084 * |
| HOMA F L ET AL: "Isolation and expression of a cDNA clone encoding a bovine UDP-GalNAc: polypeptide N-acetylgalactosaminyltransferase" JOURNAL OF BIOLOGICAL CHEMISTRY, AMERICAN SOCIETY OF BIOLOCHEMICAL BIOLOGISTS, BIRMINGHAM,; US, vol. 268, no. 17, 15 June 1993 (1993-06-15), pages 12609-12616, XP002965844 ISSN: 0021-9258 * |
| KUROSAWA N ET AL: "Molecular cloning and expression of GalNAc alpha2,6-Sialyltransferase" JOURNAL OF BIOLOGICAL CHEMISTRY, AMERICAN SOCIETY OF BIOLOCHEMICAL BIOLOGISTS, BIRMINGHAM,; US, vol. 269, no. 2, 14 January 1994 (1994-01-14), pages 1402-1409, XP002486638 ISSN: 0021-9258 * |
| See also references of WO2005121331A2 * |
| WANDALL H H ET AL: "Substrate specificities of three members of the human UPD-N-acetyl-alpha-D-galactosamine : Polypeptide N-acetylgalactosaminyltransferase family, GalNAc-T1, -T2, and -T3" JOURNAL OF BIOLOGICAL CHEMISTRY, AMERICAN SOCIETY OF BIOLOCHEMICAL BIOLOGISTS, BIRMINGHAM,; US, vol. 272, no. 38, 19 September 1997 (1997-09-19), pages 23503-23514, XP002904621 ISSN: 0021-9258 * |
Also Published As
| Publication number | Publication date |
|---|---|
| EP1765992A4 (de) | 2009-01-07 |
| WO2005121331A8 (en) | 2006-03-09 |
| JP2008512085A (ja) | 2008-04-24 |
| WO2005121331A9 (en) | 2006-06-01 |
| WO2005121331A2 (en) | 2005-12-22 |
| WO2005121331A3 (en) | 2007-08-30 |
Similar Documents
| Publication | Publication Date | Title |
|---|---|---|
| WO2005121331A2 (en) | Truncated galnact2 polypeptides and nucleic acids | |
| US20080206810A1 (en) | Truncated St6galnaci Polypeptides and Nucleic Acids | |
| US9809835B2 (en) | Quantitative control of sialylation | |
| EP1539989A2 (de) | Synthese von glycoproteinen unter verwendung bakterieller glycosyltransferasen | |
| JP6511045B2 (ja) | N末端が切り詰められたベータ−ガラクトシドアルファ−2,6−シアリルトランスフェラーゼ変異体を用いた糖タンパク質のモノ−およびバイ−シアリル化のためのプロセス | |
| CN105431530A (zh) | N末端截短的糖基转移酶 | |
| JP2011167200A (ja) | H.pyloriフコシルトランスフェラーゼ | |
| US20170204381A1 (en) | Pmst1 mutants for chemoenzymatic synthesis of sialyl lewis x compounds | |
| CA2555109C (en) | Methods of refolding mammalian glycosyltransferases | |
| US9783838B2 (en) | PmST3 enzyme for chemoenzymatic synthesis of alpha-2-3-sialosides | |
| US9938510B2 (en) | Photobacterium sp. alpha-2-6-sialyltransferase variants | |
| CN101151367B (zh) | 哺乳动物糖基转移酶的重折叠方法 | |
| WO2008097829A2 (en) | Large scale production of eukaryotic n-acetylglucosaminyltransferase i in bacteria | |
| US20090047710A1 (en) | ST3Gal-1/ST6GalNAc-1 Chimeras | |
| JPWO2012014980A1 (ja) | 新規酵素タンパク質、当該酵素タンパク質の製造方法及び当該酵素タンパク質をコードする遺伝子 | |
| Lattard et al. | Purification and characterization of a soluble form of the recombinant human galactose-β1, 3-glucuronosyltransferase I expressed in the yeast Pichia pastoris | |
| WO2024097788A1 (en) | Glycosyltransferase engineering for chemoenzymatic total synthesis of gangliosides |
Legal Events
| Date | Code | Title | Description |
|---|---|---|---|
| PUAI | Public reference made under article 153(3) epc to a published international application that has entered the european phase |
Free format text: ORIGINAL CODE: 0009012 |
|
| 17P | Request for examination filed |
Effective date: 20061227 |
|
| AK | Designated contracting states |
Kind code of ref document: A2 Designated state(s): AT BE BG CH CY CZ DE DK EE ES FI FR GB GR HU IE IS IT LI LT LU MC NL PL PT RO SE SI SK TR |
|
| AX | Request for extension of the european patent |
Extension state: AL BA HR LV MK YU |
|
| PUAK | Availability of information related to the publication of the international search report |
Free format text: ORIGINAL CODE: 0009015 |
|
| RIN1 | Information on inventor provided before grant (corrected) |
Inventor name: SARIBAS, SAMI Inventor name: TAUDTE, SUSANN Inventor name: CHEN, XI Inventor name: JOHNSON, KARL F. |
|
| RIC1 | Information provided on ipc code assigned before grant |
Ipc: C12N 5/10 20060101ALI20071109BHEP Ipc: C12N 15/00 20060101ALI20071109BHEP Ipc: C12N 1/20 20060101ALI20071109BHEP Ipc: C12P 21/06 20060101ALI20071109BHEP Ipc: C12P 19/18 20060101AFI20071109BHEP |
|
| RAP1 | Party data changed (applicant data changed or rights of an application transferred) |
Owner name: NEOSE TECHNOLOGIES, INC. |
|
| A4 | Supplementary search report drawn up and despatched |
Effective date: 20081209 |
|
| RIC1 | Information provided on ipc code assigned before grant |
Ipc: C12N 15/54 20060101ALI20081128BHEP Ipc: C12N 9/10 20060101AFI20081128BHEP |
|
| RAP1 | Party data changed (applicant data changed or rights of an application transferred) |
Owner name: NOVO NORDISK A/S |
|
| 17Q | First examination report despatched |
Effective date: 20120307 |
|
| STAA | Information on the status of an ep patent application or granted ep patent |
Free format text: STATUS: THE APPLICATION IS DEEMED TO BE WITHDRAWN |
|
| 18D | Application deemed to be withdrawn |
Effective date: 20120718 |