JP2013529083A - 新規のdna結合タンパク質及びその使用 - Google Patents
新規のdna結合タンパク質及びその使用 Download PDFInfo
- Publication number
- JP2013529083A JP2013529083A JP2013511148A JP2013511148A JP2013529083A JP 2013529083 A JP2013529083 A JP 2013529083A JP 2013511148 A JP2013511148 A JP 2013511148A JP 2013511148 A JP2013511148 A JP 2013511148A JP 2013529083 A JP2013529083 A JP 2013529083A
- Authority
- JP
- Japan
- Prior art keywords
- tale
- domain
- sequence
- cells
- protein
- Prior art date
- Legal status (The legal status is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the status listed.)
- Granted
Links
- 102000052510 DNA-Binding Proteins Human genes 0.000 title description 14
- 101710096438 DNA-binding protein Proteins 0.000 title description 8
- 238000000034 method Methods 0.000 claims abstract description 162
- 230000004568 DNA-binding Effects 0.000 claims abstract description 127
- 108090000765 processed proteins & peptides Proteins 0.000 claims abstract description 84
- 102000004196 processed proteins & peptides Human genes 0.000 claims abstract description 71
- 229920001184 polypeptide Polymers 0.000 claims abstract description 68
- 102000040430 polynucleotide Human genes 0.000 claims abstract description 51
- 108091033319 polynucleotide Proteins 0.000 claims abstract description 51
- 239000002157 polynucleotide Substances 0.000 claims abstract description 51
- 108090000623 proteins and genes Proteins 0.000 claims description 472
- 210000004027 cell Anatomy 0.000 claims description 308
- 102000004169 proteins and genes Human genes 0.000 claims description 258
- 238000003776 cleavage reaction Methods 0.000 claims description 189
- 230000007017 scission Effects 0.000 claims description 187
- 108020001507 fusion proteins Proteins 0.000 claims description 149
- 102000037865 fusion proteins Human genes 0.000 claims description 149
- 101710163270 Nuclease Proteins 0.000 claims description 114
- 150000007523 nucleic acids Chemical class 0.000 claims description 106
- 230000014509 gene expression Effects 0.000 claims description 103
- 102000039446 nucleic acids Human genes 0.000 claims description 100
- 108020004707 nucleic acids Proteins 0.000 claims description 100
- 150000001413 amino acids Chemical class 0.000 claims description 68
- 230000004913 activation Effects 0.000 claims description 57
- 230000006780 non-homologous end joining Effects 0.000 claims description 39
- 238000012217 deletion Methods 0.000 claims description 29
- 230000037430 deletion Effects 0.000 claims description 29
- 230000001105 regulatory effect Effects 0.000 claims description 27
- 230000001404 mediated effect Effects 0.000 claims description 23
- 239000012634 fragment Substances 0.000 claims description 22
- 238000002744 homologous recombination Methods 0.000 claims description 17
- 230000006801 homologous recombination Effects 0.000 claims description 17
- 108010042407 Endonucleases Proteins 0.000 claims description 12
- 230000002779 inactivation Effects 0.000 claims description 10
- 230000003252 repetitive effect Effects 0.000 claims description 8
- 241000251468 Actinopterygii Species 0.000 claims description 6
- 102000004533 Endonucleases Human genes 0.000 claims description 6
- 210000004102 animal cell Anatomy 0.000 claims description 6
- 108091006106 transcriptional activators Proteins 0.000 claims description 6
- 108091006107 transcriptional repressors Proteins 0.000 claims description 5
- 239000008194 pharmaceutical composition Substances 0.000 claims description 4
- 210000005253 yeast cell Anatomy 0.000 claims description 4
- 210000003527 eukaryotic cell Anatomy 0.000 claims description 3
- 230000001629 suppression Effects 0.000 claims description 2
- 102100025169 Max-binding protein MNT Human genes 0.000 claims 1
- 238000010362 genome editing Methods 0.000 abstract description 22
- 230000014493 regulation of gene expression Effects 0.000 abstract description 2
- 235000018102 proteins Nutrition 0.000 description 248
- 108020004414 DNA Proteins 0.000 description 152
- 230000004927 fusion Effects 0.000 description 127
- 241000196324 Embryophyta Species 0.000 description 94
- 230000000694 effects Effects 0.000 description 94
- 230000027455 binding Effects 0.000 description 90
- 101710149870 C-C chemokine receptor type 5 Proteins 0.000 description 88
- 102100035875 C-C chemokine receptor type 5 Human genes 0.000 description 88
- 239000013598 vector Substances 0.000 description 80
- 235000001014 amino acid Nutrition 0.000 description 70
- 238000003556 assay Methods 0.000 description 70
- 210000004899 c-terminal region Anatomy 0.000 description 70
- 229940024606 amino acid Drugs 0.000 description 67
- 239000002773 nucleotide Substances 0.000 description 64
- 125000003729 nucleotide group Chemical group 0.000 description 64
- 102100029268 Neurotrophin-3 Human genes 0.000 description 53
- 239000000203 mixture Substances 0.000 description 53
- 101000634196 Homo sapiens Neurotrophin-3 Proteins 0.000 description 52
- 241001465754 Metazoa Species 0.000 description 49
- 230000006870 function Effects 0.000 description 46
- 238000013456 study Methods 0.000 description 40
- 230000032965 negative regulation of cell volume Effects 0.000 description 39
- 239000000047 product Substances 0.000 description 39
- 241000700159 Rattus Species 0.000 description 38
- 108010068250 Herpes Simplex Virus Protein Vmw65 Proteins 0.000 description 37
- 230000005782 double-strand break Effects 0.000 description 34
- 230000010354 integration Effects 0.000 description 33
- 239000013612 plasmid Substances 0.000 description 33
- 108091034117 Oligonucleotide Proteins 0.000 description 32
- 238000010459 TALEN Methods 0.000 description 32
- 238000002474 experimental method Methods 0.000 description 32
- 108700008625 Reporter Genes Proteins 0.000 description 31
- 108010043645 Transcription Activator-Like Effector Nucleases Proteins 0.000 description 31
- 210000004962 mammalian cell Anatomy 0.000 description 31
- 230000035772 mutation Effects 0.000 description 28
- 108091028043 Nucleic acid sequence Proteins 0.000 description 27
- 210000000130 stem cell Anatomy 0.000 description 27
- 108020004999 messenger RNA Proteins 0.000 description 26
- 102000040945 Transcription factor Human genes 0.000 description 25
- 108091023040 Transcription factor Proteins 0.000 description 25
- 230000001413 cellular effect Effects 0.000 description 24
- 230000008685 targeting Effects 0.000 description 24
- 125000003275 alpha amino acid group Chemical group 0.000 description 23
- 230000009261 transgenic effect Effects 0.000 description 23
- 108010077544 Chromatin Proteins 0.000 description 22
- 108091026890 Coding region Proteins 0.000 description 22
- 210000003483 chromatin Anatomy 0.000 description 22
- 239000005090 green fluorescent protein Substances 0.000 description 21
- 238000004458 analytical method Methods 0.000 description 20
- 238000013461 design Methods 0.000 description 20
- 210000002257 embryonic structure Anatomy 0.000 description 20
- 238000003780 insertion Methods 0.000 description 20
- 230000037431 insertion Effects 0.000 description 20
- 210000001519 tissue Anatomy 0.000 description 20
- 102000004190 Enzymes Human genes 0.000 description 19
- 108090000790 Enzymes Proteins 0.000 description 19
- 238000001415 gene therapy Methods 0.000 description 19
- 238000000338 in vitro Methods 0.000 description 19
- -1 IL2Rγ Proteins 0.000 description 18
- 108700019146 Transgenes Proteins 0.000 description 18
- 238000001727 in vivo Methods 0.000 description 18
- 238000004519 manufacturing process Methods 0.000 description 18
- 241000589634 Xanthomonas Species 0.000 description 17
- 230000003612 virological effect Effects 0.000 description 17
- 238000002965 ELISA Methods 0.000 description 16
- 108091030071 RNAI Proteins 0.000 description 16
- 238000006243 chemical reaction Methods 0.000 description 16
- 230000009368 gene silencing by RNA Effects 0.000 description 16
- 239000013603 viral vector Substances 0.000 description 16
- 101001000998 Homo sapiens Protein phosphatase 1 regulatory subunit 12C Proteins 0.000 description 15
- 102100034349 Integrase Human genes 0.000 description 15
- 108010061833 Integrases Proteins 0.000 description 15
- 102100035620 Protein phosphatase 1 regulatory subunit 12C Human genes 0.000 description 15
- 239000013604 expression vector Substances 0.000 description 15
- 230000008439 repair process Effects 0.000 description 15
- 108091008146 restriction endonucleases Proteins 0.000 description 15
- 238000006467 substitution reaction Methods 0.000 description 15
- 241000232299 Ralstonia Species 0.000 description 14
- 230000003321 amplification Effects 0.000 description 14
- 230000003993 interaction Effects 0.000 description 14
- 238000003199 nucleic acid amplification method Methods 0.000 description 14
- 230000008569 process Effects 0.000 description 14
- 230000007018 DNA scission Effects 0.000 description 13
- 102000018120 Recombinases Human genes 0.000 description 13
- 108010091086 Recombinases Proteins 0.000 description 13
- 101710185494 Zinc finger protein Proteins 0.000 description 13
- 102100023597 Zinc finger protein 816 Human genes 0.000 description 13
- 238000013459 approach Methods 0.000 description 13
- 238000012163 sequencing technique Methods 0.000 description 13
- 241000700605 Viruses Species 0.000 description 12
- 238000011161 development Methods 0.000 description 12
- 230000018109 developmental process Effects 0.000 description 12
- 210000001671 embryonic stem cell Anatomy 0.000 description 12
- 239000000499 gel Substances 0.000 description 12
- RXWNCPJZOCPEPQ-NVWDDTSBSA-N puromycin Chemical compound C1=CC(OC)=CC=C1C[C@H](N)C(=O)N[C@H]1[C@@H](O)[C@H](N2C3=NC=NC(=C3N=C2)N(C)C)O[C@@H]1CO RXWNCPJZOCPEPQ-NVWDDTSBSA-N 0.000 description 12
- 230000001177 retroviral effect Effects 0.000 description 12
- 239000000523 sample Substances 0.000 description 12
- 238000012360 testing method Methods 0.000 description 12
- 238000001890 transfection Methods 0.000 description 12
- 238000012546 transfer Methods 0.000 description 12
- 102000053602 DNA Human genes 0.000 description 11
- 108060001084 Luciferase Proteins 0.000 description 11
- 206010028980 Neoplasm Diseases 0.000 description 11
- 238000012408 PCR amplification Methods 0.000 description 11
- 240000008042 Zea mays Species 0.000 description 11
- 235000002017 Zea mays subsp mays Nutrition 0.000 description 11
- HCHKCACWOHOZIP-UHFFFAOYSA-N Zinc Chemical compound [Zn] HCHKCACWOHOZIP-UHFFFAOYSA-N 0.000 description 11
- 210000000349 chromosome Anatomy 0.000 description 11
- 238000010367 cloning Methods 0.000 description 11
- UYTPUPDQBNUYGX-UHFFFAOYSA-N guanine Chemical class O=C1NC(N)=NC2=C1N=CN2 UYTPUPDQBNUYGX-UHFFFAOYSA-N 0.000 description 11
- 230000004048 modification Effects 0.000 description 11
- 238000012986 modification Methods 0.000 description 11
- 239000003921 oil Substances 0.000 description 11
- 235000019198 oils Nutrition 0.000 description 11
- 150000003384 small molecules Chemical class 0.000 description 11
- 230000001225 therapeutic effect Effects 0.000 description 11
- 241000701161 unidentified adenovirus Species 0.000 description 11
- 229910052725 zinc Inorganic materials 0.000 description 11
- 239000011701 zinc Substances 0.000 description 11
- 241000702421 Dependoparvovirus Species 0.000 description 10
- ROHFNLRQFUQHCH-YFKPBYRVSA-N L-leucine Chemical compound CC(C)C[C@H](N)C(O)=O ROHFNLRQFUQHCH-YFKPBYRVSA-N 0.000 description 10
- 239000005089 Luciferase Substances 0.000 description 10
- 108010019530 Vascular Endothelial Growth Factors Proteins 0.000 description 10
- 230000008859 change Effects 0.000 description 10
- 230000002759 chromosomal effect Effects 0.000 description 10
- 239000000539 dimer Substances 0.000 description 10
- 208000037265 diseases, disorders, signs and symptoms Diseases 0.000 description 10
- 238000012239 gene modification Methods 0.000 description 10
- 238000002347 injection Methods 0.000 description 10
- 239000007924 injection Substances 0.000 description 10
- 239000006166 lysate Substances 0.000 description 10
- 230000006798 recombination Effects 0.000 description 10
- 238000005215 recombination Methods 0.000 description 10
- 241000894007 species Species 0.000 description 10
- ROHFNLRQFUQHCH-UHFFFAOYSA-N Leucine Natural products CC(C)CC(N)C(O)=O ROHFNLRQFUQHCH-UHFFFAOYSA-N 0.000 description 9
- KDXKERNSBIXSRK-UHFFFAOYSA-N Lysine Natural products NCCCCC(N)C(O)=O KDXKERNSBIXSRK-UHFFFAOYSA-N 0.000 description 9
- 108090000742 Neurotrophin 3 Proteins 0.000 description 9
- 108010020764 Transposases Proteins 0.000 description 9
- 102000008579 Transposases Human genes 0.000 description 9
- 102000005789 Vascular Endothelial Growth Factors Human genes 0.000 description 9
- 235000005824 Zea mays ssp. parviglumis Nutrition 0.000 description 9
- 238000010276 construction Methods 0.000 description 9
- 235000005822 corn Nutrition 0.000 description 9
- 230000005017 genetic modification Effects 0.000 description 9
- 235000013617 genetically modified food Nutrition 0.000 description 9
- 230000001939 inductive effect Effects 0.000 description 9
- 210000001161 mammalian embryo Anatomy 0.000 description 9
- 238000004806 packaging method and process Methods 0.000 description 9
- 230000012743 protein tagging Effects 0.000 description 9
- 238000013518 transcription Methods 0.000 description 9
- 230000035897 transcription Effects 0.000 description 9
- 230000002103 transcriptional effect Effects 0.000 description 9
- 108020004705 Codon Proteins 0.000 description 8
- 241000233866 Fungi Species 0.000 description 8
- 235000010469 Glycine max Nutrition 0.000 description 8
- 125000000539 amino acid group Chemical group 0.000 description 8
- 201000010099 disease Diseases 0.000 description 8
- 230000002068 genetic effect Effects 0.000 description 8
- 239000003446 ligand Substances 0.000 description 8
- 239000002502 liposome Substances 0.000 description 8
- 239000002245 particle Substances 0.000 description 8
- 229950010131 puromycin Drugs 0.000 description 8
- 238000011144 upstream manufacturing Methods 0.000 description 8
- 108700028369 Alleles Proteins 0.000 description 7
- WHUUTDBJXJRKMK-UHFFFAOYSA-N Glutamic acid Natural products OC(=O)C(N)CCC(O)=O WHUUTDBJXJRKMK-UHFFFAOYSA-N 0.000 description 7
- 244000068988 Glycine max Species 0.000 description 7
- 241000254158 Lampyridae Species 0.000 description 7
- 102000055027 Protein Methyltransferases Human genes 0.000 description 7
- 108700040121 Protein Methyltransferases Proteins 0.000 description 7
- JLCPHMBAVCMARE-UHFFFAOYSA-N [3-[[3-[[3-[[3-[[3-[[3-[[3-[[3-[[3-[[3-[[3-[[5-(2-amino-6-oxo-1H-purin-9-yl)-3-[[3-[[3-[[3-[[3-[[3-[[5-(2-amino-6-oxo-1H-purin-9-yl)-3-[[5-(2-amino-6-oxo-1H-purin-9-yl)-3-hydroxyoxolan-2-yl]methoxy-hydroxyphosphoryl]oxyoxolan-2-yl]methoxy-hydroxyphosphoryl]oxy-5-(5-methyl-2,4-dioxopyrimidin-1-yl)oxolan-2-yl]methoxy-hydroxyphosphoryl]oxy-5-(6-aminopurin-9-yl)oxolan-2-yl]methoxy-hydroxyphosphoryl]oxy-5-(6-aminopurin-9-yl)oxolan-2-yl]methoxy-hydroxyphosphoryl]oxy-5-(6-aminopurin-9-yl)oxolan-2-yl]methoxy-hydroxyphosphoryl]oxy-5-(6-aminopurin-9-yl)oxolan-2-yl]methoxy-hydroxyphosphoryl]oxyoxolan-2-yl]methoxy-hydroxyphosphoryl]oxy-5-(5-methyl-2,4-dioxopyrimidin-1-yl)oxolan-2-yl]methoxy-hydroxyphosphoryl]oxy-5-(4-amino-2-oxopyrimidin-1-yl)oxolan-2-yl]methoxy-hydroxyphosphoryl]oxy-5-(5-methyl-2,4-dioxopyrimidin-1-yl)oxolan-2-yl]methoxy-hydroxyphosphoryl]oxy-5-(5-methyl-2,4-dioxopyrimidin-1-yl)oxolan-2-yl]methoxy-hydroxyphosphoryl]oxy-5-(6-aminopurin-9-yl)oxolan-2-yl]methoxy-hydroxyphosphoryl]oxy-5-(6-aminopurin-9-yl)oxolan-2-yl]methoxy-hydroxyphosphoryl]oxy-5-(4-amino-2-oxopyrimidin-1-yl)oxolan-2-yl]methoxy-hydroxyphosphoryl]oxy-5-(4-amino-2-oxopyrimidin-1-yl)oxolan-2-yl]methoxy-hydroxyphosphoryl]oxy-5-(4-amino-2-oxopyrimidin-1-yl)oxolan-2-yl]methoxy-hydroxyphosphoryl]oxy-5-(6-aminopurin-9-yl)oxolan-2-yl]methoxy-hydroxyphosphoryl]oxy-5-(4-amino-2-oxopyrimidin-1-yl)oxolan-2-yl]methyl [5-(6-aminopurin-9-yl)-2-(hydroxymethyl)oxolan-3-yl] hydrogen phosphate Polymers Cc1cn(C2CC(OP(O)(=O)OCC3OC(CC3OP(O)(=O)OCC3OC(CC3O)n3cnc4c3nc(N)[nH]c4=O)n3cnc4c3nc(N)[nH]c4=O)C(COP(O)(=O)OC3CC(OC3COP(O)(=O)OC3CC(OC3COP(O)(=O)OC3CC(OC3COP(O)(=O)OC3CC(OC3COP(O)(=O)OC3CC(OC3COP(O)(=O)OC3CC(OC3COP(O)(=O)OC3CC(OC3COP(O)(=O)OC3CC(OC3COP(O)(=O)OC3CC(OC3COP(O)(=O)OC3CC(OC3COP(O)(=O)OC3CC(OC3COP(O)(=O)OC3CC(OC3COP(O)(=O)OC3CC(OC3COP(O)(=O)OC3CC(OC3COP(O)(=O)OC3CC(OC3COP(O)(=O)OC3CC(OC3COP(O)(=O)OC3CC(OC3CO)n3cnc4c(N)ncnc34)n3ccc(N)nc3=O)n3cnc4c(N)ncnc34)n3ccc(N)nc3=O)n3ccc(N)nc3=O)n3ccc(N)nc3=O)n3cnc4c(N)ncnc34)n3cnc4c(N)ncnc34)n3cc(C)c(=O)[nH]c3=O)n3cc(C)c(=O)[nH]c3=O)n3ccc(N)nc3=O)n3cc(C)c(=O)[nH]c3=O)n3cnc4c3nc(N)[nH]c4=O)n3cnc4c(N)ncnc34)n3cnc4c(N)ncnc34)n3cnc4c(N)ncnc34)n3cnc4c(N)ncnc34)O2)c(=O)[nH]c1=O JLCPHMBAVCMARE-UHFFFAOYSA-N 0.000 description 7
- 239000012190 activator Substances 0.000 description 7
- 238000007792 addition Methods 0.000 description 7
- 238000004520 electroporation Methods 0.000 description 7
- 210000003958 hematopoietic stem cell Anatomy 0.000 description 7
- 238000005259 measurement Methods 0.000 description 7
- 230000037361 pathway Effects 0.000 description 7
- 230000009466 transformation Effects 0.000 description 7
- 241000894006 Bacteria Species 0.000 description 6
- 102000014914 Carrier Proteins Human genes 0.000 description 6
- 108700020911 DNA-Binding Proteins Proteins 0.000 description 6
- 102100031780 Endonuclease Human genes 0.000 description 6
- 108090000331 Firefly luciferases Proteins 0.000 description 6
- 108700011259 MicroRNAs Proteins 0.000 description 6
- 241000699670 Mus sp. Species 0.000 description 6
- 108010077850 Nuclear Localization Signals Proteins 0.000 description 6
- 102100035423 POU domain, class 5, transcription factor 1 Human genes 0.000 description 6
- 101710126211 POU domain, class 5, transcription factor 1 Proteins 0.000 description 6
- 108091027967 Small hairpin RNA Proteins 0.000 description 6
- 108010017070 Zinc Finger Nucleases Proteins 0.000 description 6
- 108091008324 binding proteins Proteins 0.000 description 6
- 230000015572 biosynthetic process Effects 0.000 description 6
- OPTASPLRGRRNAP-UHFFFAOYSA-N cytosine Chemical compound NC=1C=CNC(=O)N=1 OPTASPLRGRRNAP-UHFFFAOYSA-N 0.000 description 6
- 230000007423 decrease Effects 0.000 description 6
- 235000014113 dietary fatty acids Nutrition 0.000 description 6
- 230000029087 digestion Effects 0.000 description 6
- 230000034431 double-strand break repair via homologous recombination Effects 0.000 description 6
- 239000003814 drug Substances 0.000 description 6
- 238000005516 engineering process Methods 0.000 description 6
- 108010048367 enhanced green fluorescent protein Proteins 0.000 description 6
- 229930195729 fatty acid Natural products 0.000 description 6
- 239000000194 fatty acid Substances 0.000 description 6
- 150000004665 fatty acids Chemical class 0.000 description 6
- 238000010353 genetic engineering Methods 0.000 description 6
- 239000000833 heterodimer Substances 0.000 description 6
- 238000009396 hybridization Methods 0.000 description 6
- 150000002632 lipids Chemical class 0.000 description 6
- 239000003550 marker Substances 0.000 description 6
- 239000002679 microRNA Substances 0.000 description 6
- 230000029279 positive regulation of transcription, DNA-dependent Effects 0.000 description 6
- 108020001580 protein domains Proteins 0.000 description 6
- 230000022532 regulation of transcription, DNA-dependent Effects 0.000 description 6
- 230000035939 shock Effects 0.000 description 6
- 239000004055 small Interfering RNA Substances 0.000 description 6
- 125000006850 spacer group Chemical group 0.000 description 6
- 238000010561 standard procedure Methods 0.000 description 6
- 239000000126 substance Substances 0.000 description 6
- RWQNBRDOKXIBIV-UHFFFAOYSA-N thymine Chemical compound CC1=CNC(=O)NC1=O RWQNBRDOKXIBIV-UHFFFAOYSA-N 0.000 description 6
- 239000013638 trimer Substances 0.000 description 6
- WRIDQFICGBMAFQ-UHFFFAOYSA-N (E)-8-Octadecenoic acid Natural products CCCCCCCCCC=CCCCCCCC(O)=O WRIDQFICGBMAFQ-UHFFFAOYSA-N 0.000 description 5
- LQJBNNIYVWPHFW-UHFFFAOYSA-N 20:1omega9c fatty acid Natural products CCCCCCCCCCC=CCCCCCCCC(O)=O LQJBNNIYVWPHFW-UHFFFAOYSA-N 0.000 description 5
- QSBYPNXLFMSGKH-UHFFFAOYSA-N 9-Heptadecensaeure Natural products CCCCCCCC=CCCCCCCCC(O)=O QSBYPNXLFMSGKH-UHFFFAOYSA-N 0.000 description 5
- 108010016626 Dipeptides Proteins 0.000 description 5
- 101000595674 Homo sapiens Pituitary homeobox 3 Proteins 0.000 description 5
- 239000005642 Oleic acid Substances 0.000 description 5
- ZQPPMHVWECSIRJ-UHFFFAOYSA-N Oleic acid Natural products CCCCCCCCC=CCCCCCCCC(O)=O ZQPPMHVWECSIRJ-UHFFFAOYSA-N 0.000 description 5
- 241000283973 Oryctolagus cuniculus Species 0.000 description 5
- 102100035100 Transcription factor p65 Human genes 0.000 description 5
- 238000000137 annealing Methods 0.000 description 5
- RIOXQFHNBCKOKP-UHFFFAOYSA-N benomyl Chemical compound C1=CC=C2N(C(=O)NCCCC)C(NC(=O)OC)=NC2=C1 RIOXQFHNBCKOKP-UHFFFAOYSA-N 0.000 description 5
- MITFXPHMIHQXPI-UHFFFAOYSA-N benzoxaprofen Natural products N=1C2=CC(C(C(O)=O)C)=CC=C2OC=1C1=CC=C(Cl)C=C1 MITFXPHMIHQXPI-UHFFFAOYSA-N 0.000 description 5
- 230000033228 biological regulation Effects 0.000 description 5
- 210000004369 blood Anatomy 0.000 description 5
- 239000008280 blood Substances 0.000 description 5
- 239000002299 complementary DNA Substances 0.000 description 5
- 238000010586 diagram Methods 0.000 description 5
- QXJSBBXBKPUZAA-UHFFFAOYSA-N isooleic acid Natural products CCCCCCCC=CCCCCCCCCC(O)=O QXJSBBXBKPUZAA-UHFFFAOYSA-N 0.000 description 5
- 238000000520 microinjection Methods 0.000 description 5
- 239000003607 modifier Substances 0.000 description 5
- ZQPPMHVWECSIRJ-KTKRTIGZSA-N oleic acid Chemical compound CCCCCCCC\C=C/CCCCCCCC(O)=O ZQPPMHVWECSIRJ-KTKRTIGZSA-N 0.000 description 5
- 210000001778 pluripotent stem cell Anatomy 0.000 description 5
- 239000000243 solution Substances 0.000 description 5
- 238000010361 transduction Methods 0.000 description 5
- 230000026683 transduction Effects 0.000 description 5
- 230000014616 translation Effects 0.000 description 5
- 230000004614 tumor growth Effects 0.000 description 5
- 239000003981 vehicle Substances 0.000 description 5
- 229930024421 Adenine Natural products 0.000 description 4
- GFFGJBXGBJISGV-UHFFFAOYSA-N Adenine Chemical compound NC1=NC=NC2=C1N=CN2 GFFGJBXGBJISGV-UHFFFAOYSA-N 0.000 description 4
- 241000195493 Cryptophyta Species 0.000 description 4
- 108090000695 Cytokines Proteins 0.000 description 4
- 102000004127 Cytokines Human genes 0.000 description 4
- 230000033616 DNA repair Effects 0.000 description 4
- QNAYBMKLOCPYGJ-REOHCLBHSA-N L-alanine Chemical compound C[C@H](N)C(O)=O QNAYBMKLOCPYGJ-REOHCLBHSA-N 0.000 description 4
- 101150036143 NTF3 gene Proteins 0.000 description 4
- 206010029113 Neovascularisation Diseases 0.000 description 4
- 102000001253 Protein Kinase Human genes 0.000 description 4
- 108010052090 Renilla Luciferases Proteins 0.000 description 4
- 108091081062 Repeated sequence (DNA) Proteins 0.000 description 4
- 240000004808 Saccharomyces cerevisiae Species 0.000 description 4
- 238000002105 Southern blotting Methods 0.000 description 4
- 210000001744 T-lymphocyte Anatomy 0.000 description 4
- 229960000643 adenine Drugs 0.000 description 4
- 239000011543 agarose gel Substances 0.000 description 4
- 235000004279 alanine Nutrition 0.000 description 4
- 230000004075 alteration Effects 0.000 description 4
- 230000001580 bacterial effect Effects 0.000 description 4
- 230000008901 benefit Effects 0.000 description 4
- 239000011230 binding agent Substances 0.000 description 4
- 210000002459 blastocyst Anatomy 0.000 description 4
- 201000011510 cancer Diseases 0.000 description 4
- 210000004978 chinese hamster ovary cell Anatomy 0.000 description 4
- 238000007796 conventional method Methods 0.000 description 4
- 230000001419 dependent effect Effects 0.000 description 4
- ZMMJGEGLRURXTF-UHFFFAOYSA-N ethidium bromide Chemical compound [Br-].C12=CC(N)=CC=C2C2=CC=C(N)C=C2[N+](CC)=C1C1=CC=CC=C1 ZMMJGEGLRURXTF-UHFFFAOYSA-N 0.000 description 4
- 229960005542 ethidium bromide Drugs 0.000 description 4
- 238000009472 formulation Methods 0.000 description 4
- 230000005714 functional activity Effects 0.000 description 4
- 210000004602 germ cell Anatomy 0.000 description 4
- 238000000099 in vitro assay Methods 0.000 description 4
- 238000001638 lipofection Methods 0.000 description 4
- 238000013507 mapping Methods 0.000 description 4
- 239000011159 matrix material Substances 0.000 description 4
- 210000002901 mesenchymal stem cell Anatomy 0.000 description 4
- 210000004897 n-terminal region Anatomy 0.000 description 4
- 210000001178 neural stem cell Anatomy 0.000 description 4
- 230000030648 nucleus localization Effects 0.000 description 4
- 238000005457 optimization Methods 0.000 description 4
- 230000002018 overexpression Effects 0.000 description 4
- 230000036961 partial effect Effects 0.000 description 4
- 230000001717 pathogenic effect Effects 0.000 description 4
- 210000003819 peripheral blood mononuclear cell Anatomy 0.000 description 4
- 238000002823 phage display Methods 0.000 description 4
- 229920000642 polymer Polymers 0.000 description 4
- 125000002924 primary amino group Chemical group [H]N([H])* 0.000 description 4
- 108060006633 protein kinase Proteins 0.000 description 4
- 230000002829 reductive effect Effects 0.000 description 4
- 230000010076 replication Effects 0.000 description 4
- 238000011160 research Methods 0.000 description 4
- 230000004044 response Effects 0.000 description 4
- 238000012552 review Methods 0.000 description 4
- 238000013519 translation Methods 0.000 description 4
- DCXXMTOCNZCJGO-UHFFFAOYSA-N tristearoylglycerol Chemical compound CCCCCCCCCCCCCCCCCC(=O)OCC(OC(=O)CCCCCCCCCCCCCCCCC)COC(=O)CCCCCCCCCCCCCCCCC DCXXMTOCNZCJGO-UHFFFAOYSA-N 0.000 description 4
- 230000009452 underexpressoin Effects 0.000 description 4
- 241001430294 unidentified retrovirus Species 0.000 description 4
- 210000004291 uterus Anatomy 0.000 description 4
- 239000013607 AAV vector Substances 0.000 description 3
- 241000589158 Agrobacterium Species 0.000 description 3
- 241000589155 Agrobacterium tumefaciens Species 0.000 description 3
- 101100517192 Arabidopsis thaliana NRPD1 gene Proteins 0.000 description 3
- 101100038200 Arabidopsis thaliana RPD1 gene Proteins 0.000 description 3
- IJGRMHOSHXDMSA-UHFFFAOYSA-N Atomic nitrogen Chemical compound N#N IJGRMHOSHXDMSA-UHFFFAOYSA-N 0.000 description 3
- 241000283690 Bos taurus Species 0.000 description 3
- 235000004977 Brassica sinapistrum Nutrition 0.000 description 3
- 101150017501 CCR5 gene Proteins 0.000 description 3
- 241000244203 Caenorhabditis elegans Species 0.000 description 3
- 101100297347 Caenorhabditis elegans pgl-3 gene Proteins 0.000 description 3
- 229920000742 Cotton Polymers 0.000 description 3
- 241000450599 DNA viruses Species 0.000 description 3
- 241000206602 Eukaryota Species 0.000 description 3
- 108091060211 Expressed sequence tag Proteins 0.000 description 3
- 101150016855 FAD2-1 gene Proteins 0.000 description 3
- 241000219146 Gossypium Species 0.000 description 3
- 244000020551 Helianthus annuus Species 0.000 description 3
- 235000003222 Helianthus annuus Nutrition 0.000 description 3
- 241000238631 Hexapoda Species 0.000 description 3
- 235000007340 Hordeum vulgare Nutrition 0.000 description 3
- 240000005979 Hordeum vulgare Species 0.000 description 3
- 241000725303 Human immunodeficiency virus Species 0.000 description 3
- 206010020649 Hyperkeratosis Diseases 0.000 description 3
- 102100021244 Integral membrane protein GPR180 Human genes 0.000 description 3
- 240000008415 Lactuca sativa Species 0.000 description 3
- 235000007688 Lycopersicon esculentum Nutrition 0.000 description 3
- 239000004472 Lysine Substances 0.000 description 3
- 241000124008 Mammalia Species 0.000 description 3
- 108060004795 Methyltransferase Proteins 0.000 description 3
- 241000699666 Mus <mouse, genus> Species 0.000 description 3
- 108091061960 Naked DNA Proteins 0.000 description 3
- 108700019961 Neoplasm Genes Proteins 0.000 description 3
- 102000048850 Neoplasm Genes Human genes 0.000 description 3
- 102000011931 Nucleoproteins Human genes 0.000 description 3
- 108010061100 Nucleoproteins Proteins 0.000 description 3
- 108010047956 Nucleosomes Proteins 0.000 description 3
- 206010033799 Paralysis Diseases 0.000 description 3
- 108700019535 Phosphoprotein Phosphatases Proteins 0.000 description 3
- 102000045595 Phosphoprotein Phosphatases Human genes 0.000 description 3
- 102100036088 Pituitary homeobox 3 Human genes 0.000 description 3
- 108010083644 Ribonucleases Proteins 0.000 description 3
- 102000006382 Ribonucleases Human genes 0.000 description 3
- 101100473190 Saccharomyces cerevisiae (strain ATCC 204508 / S288c) RPN1 gene Proteins 0.000 description 3
- 101100042631 Saccharomyces cerevisiae (strain ATCC 204508 / S288c) SIN3 gene Proteins 0.000 description 3
- 240000003768 Solanum lycopersicum Species 0.000 description 3
- 244000062793 Sorghum vulgare Species 0.000 description 3
- IQFYYKKMVGJFEH-XLPZGREQSA-N Thymidine Chemical group O=C1NC(=O)C(C)=CN1[C@@H]1O[C@H](CO)[C@@H](O)C1 IQFYYKKMVGJFEH-XLPZGREQSA-N 0.000 description 3
- 235000021307 Triticum Nutrition 0.000 description 3
- 244000098338 Triticum aestivum Species 0.000 description 3
- KZSNJWFQEVHDMF-UHFFFAOYSA-N Valine Natural products CC(C)C(N)C(O)=O KZSNJWFQEVHDMF-UHFFFAOYSA-N 0.000 description 3
- 108700005077 Viral Genes Proteins 0.000 description 3
- 238000009825 accumulation Methods 0.000 description 3
- 239000002253 acid Substances 0.000 description 3
- 238000010171 animal model Methods 0.000 description 3
- 229960001230 asparagine Drugs 0.000 description 3
- 238000013357 binding ELISA Methods 0.000 description 3
- 239000002551 biofuel Substances 0.000 description 3
- 238000009709 capacitor discharge sintering Methods 0.000 description 3
- 230000003197 catalytic effect Effects 0.000 description 3
- 108091092356 cellular DNA Proteins 0.000 description 3
- 239000003153 chemical reaction reagent Substances 0.000 description 3
- 239000003795 chemical substances by application Substances 0.000 description 3
- 150000001875 compounds Chemical class 0.000 description 3
- 210000004748 cultured cell Anatomy 0.000 description 3
- 229940104302 cytosine Drugs 0.000 description 3
- 230000002950 deficient Effects 0.000 description 3
- 230000009274 differential gene expression Effects 0.000 description 3
- 238000006471 dimerization reaction Methods 0.000 description 3
- 238000010494 dissociation reaction Methods 0.000 description 3
- 230000005593 dissociations Effects 0.000 description 3
- 230000003828 downregulation Effects 0.000 description 3
- 239000003937 drug carrier Substances 0.000 description 3
- 241001493065 dsRNA viruses Species 0.000 description 3
- 239000012636 effector Substances 0.000 description 3
- 238000011156 evaluation Methods 0.000 description 3
- 230000002538 fungal effect Effects 0.000 description 3
- 125000003630 glycyl group Chemical group [H]N([H])C([H])([H])C(*)=O 0.000 description 3
- 239000000710 homodimer Substances 0.000 description 3
- 238000011534 incubation Methods 0.000 description 3
- 238000001990 intravenous administration Methods 0.000 description 3
- 210000004185 liver Anatomy 0.000 description 3
- 125000001360 methionine group Chemical group N[C@@H](CCSC)C(=O)* 0.000 description 3
- 239000000178 monomer Substances 0.000 description 3
- 210000001665 muscle stem cell Anatomy 0.000 description 3
- 231100000350 mutagenesis Toxicity 0.000 description 3
- 210000001623 nucleosome Anatomy 0.000 description 3
- 210000004940 nucleus Anatomy 0.000 description 3
- 230000003032 phytopathogenic effect Effects 0.000 description 3
- 238000002360 preparation method Methods 0.000 description 3
- 102000005962 receptors Human genes 0.000 description 3
- 108020003175 receptors Proteins 0.000 description 3
- 238000012216 screening Methods 0.000 description 3
- 230000037432 silent mutation Effects 0.000 description 3
- 238000003860 storage Methods 0.000 description 3
- 239000000758 substrate Substances 0.000 description 3
- 229940113082 thymine Drugs 0.000 description 3
- 230000003827 upregulation Effects 0.000 description 3
- 235000015112 vegetable and seed oil Nutrition 0.000 description 3
- ZOOGRGPOEVQQDX-UUOKFMHZSA-N 3',5'-cyclic GMP Chemical compound C([C@H]1O2)OP(O)(=O)O[C@H]1[C@@H](O)[C@@H]2N1C(N=C(NC2=O)N)=C2N=C1 ZOOGRGPOEVQQDX-UUOKFMHZSA-N 0.000 description 2
- 108010013043 Acetylesterase Proteins 0.000 description 2
- 102100033647 Activity-regulated cytoskeleton-associated protein Human genes 0.000 description 2
- 101001030716 Arabidopsis thaliana Histone deacetylase HDT1 Proteins 0.000 description 2
- 235000000832 Ayote Nutrition 0.000 description 2
- 241000219310 Beta vulgaris subsp. vulgaris Species 0.000 description 2
- 240000002791 Brassica napus Species 0.000 description 2
- 125000001433 C-terminal amino-acid group Chemical group 0.000 description 2
- 241000282472 Canis lupus familiaris Species 0.000 description 2
- 235000002566 Capsicum Nutrition 0.000 description 2
- 108090000994 Catalytic RNA Proteins 0.000 description 2
- 102000053642 Catalytic RNA Human genes 0.000 description 2
- 102000008169 Co-Repressor Proteins Human genes 0.000 description 2
- 108010060434 Co-Repressor Proteins Proteins 0.000 description 2
- 108091033380 Coding strand Proteins 0.000 description 2
- 108091035707 Consensus sequence Proteins 0.000 description 2
- 235000009854 Cucurbita moschata Nutrition 0.000 description 2
- 235000009804 Cucurbita pepo subsp pepo Nutrition 0.000 description 2
- 102100036279 DNA (cytosine-5)-methyltransferase 1 Human genes 0.000 description 2
- 235000002767 Daucus carota Nutrition 0.000 description 2
- 244000000626 Daucus carota Species 0.000 description 2
- 108010008532 Deoxyribonuclease I Proteins 0.000 description 2
- 102000007260 Deoxyribonuclease I Human genes 0.000 description 2
- 238000009007 Diagnostic Kit Methods 0.000 description 2
- 208000035240 Disease Resistance Diseases 0.000 description 2
- 238000012286 ELISA Assay Methods 0.000 description 2
- 241000282326 Felis catus Species 0.000 description 2
- 241000287828 Gallus gallus Species 0.000 description 2
- 108700039691 Genetic Promoter Regions Proteins 0.000 description 2
- PEDCQBHIVMGVHV-UHFFFAOYSA-N Glycerol Natural products OCC(O)CO PEDCQBHIVMGVHV-UHFFFAOYSA-N 0.000 description 2
- 101100446349 Glycine max FAD2-1 gene Proteins 0.000 description 2
- 102100031573 Hematopoietic progenitor cell antigen CD34 Human genes 0.000 description 2
- 108091027305 Heteroduplex Proteins 0.000 description 2
- 108010036115 Histone Methyltransferases Proteins 0.000 description 2
- 102000011787 Histone Methyltransferases Human genes 0.000 description 2
- 102000003964 Histone deacetylase Human genes 0.000 description 2
- 108090000353 Histone deacetylase Proteins 0.000 description 2
- 108010033040 Histones Proteins 0.000 description 2
- 102000006947 Histones Human genes 0.000 description 2
- 241000282412 Homo Species 0.000 description 2
- 101000777663 Homo sapiens Hematopoietic progenitor cell antigen CD34 Proteins 0.000 description 2
- 101000615488 Homo sapiens Methyl-CpG-binding domain protein 2 Proteins 0.000 description 2
- 101000687346 Homo sapiens PR domain zinc finger protein 2 Proteins 0.000 description 2
- DCXYFEDJOCDNAF-REOHCLBHSA-N L-asparagine Chemical compound OC(=O)[C@@H](N)CC(N)=O DCXYFEDJOCDNAF-REOHCLBHSA-N 0.000 description 2
- CKLJMWTZIZZHCS-REOHCLBHSA-N L-aspartic acid Chemical compound OC(=O)[C@@H](N)CC(O)=O CKLJMWTZIZZHCS-REOHCLBHSA-N 0.000 description 2
- FFEARJCKVFRZRR-BYPYZUCNSA-N L-methionine Chemical compound CSCC[C@H](N)C(O)=O FFEARJCKVFRZRR-BYPYZUCNSA-N 0.000 description 2
- KZSNJWFQEVHDMF-BYPYZUCNSA-N L-valine Chemical compound CC(C)[C@H](N)C(O)=O KZSNJWFQEVHDMF-BYPYZUCNSA-N 0.000 description 2
- 235000003228 Lactuca sativa Nutrition 0.000 description 2
- 241000713666 Lentivirus Species 0.000 description 2
- 108090000364 Ligases Proteins 0.000 description 2
- 102000003960 Ligases Human genes 0.000 description 2
- 108091036060 Linker DNA Proteins 0.000 description 2
- OYHQOLUKZRVURQ-HZJYTTRNSA-N Linoleic acid Chemical compound CCCCC\C=C/C\C=C/CCCCCCCC(O)=O OYHQOLUKZRVURQ-HZJYTTRNSA-N 0.000 description 2
- 102100021299 Methyl-CpG-binding domain protein 2 Human genes 0.000 description 2
- 241000714177 Murine leukemia virus Species 0.000 description 2
- 108010057466 NF-kappa B Proteins 0.000 description 2
- 102000003945 NF-kappa B Human genes 0.000 description 2
- 238000000636 Northern blotting Methods 0.000 description 2
- 108020005497 Nuclear hormone receptor Proteins 0.000 description 2
- 102000007399 Nuclear hormone receptor Human genes 0.000 description 2
- 240000007594 Oryza sativa Species 0.000 description 2
- 235000007164 Oryza sativa Nutrition 0.000 description 2
- 102100024885 PR domain zinc finger protein 2 Human genes 0.000 description 2
- 241001494479 Pecora Species 0.000 description 2
- 102000011755 Phosphoglycerate Kinase Human genes 0.000 description 2
- 235000008331 Pinus X rigitaeda Nutrition 0.000 description 2
- 241000018646 Pinus brutia Species 0.000 description 2
- 235000011613 Pinus brutia Nutrition 0.000 description 2
- 239000002202 Polyethylene glycol Substances 0.000 description 2
- 241000125945 Protoparvovirus Species 0.000 description 2
- 108020005067 RNA Splice Sites Proteins 0.000 description 2
- 101100087805 Ralstonia solanacearum rip19 gene Proteins 0.000 description 2
- 235000007238 Secale cereale Nutrition 0.000 description 2
- 238000012300 Sequence Analysis Methods 0.000 description 2
- 241000713311 Simian immunodeficiency virus Species 0.000 description 2
- FAPWRFPIFSIZLT-UHFFFAOYSA-M Sodium chloride Chemical compound [Na+].[Cl-] FAPWRFPIFSIZLT-UHFFFAOYSA-M 0.000 description 2
- 235000002597 Solanum melongena Nutrition 0.000 description 2
- 244000061458 Solanum melongena Species 0.000 description 2
- 235000002595 Solanum tuberosum Nutrition 0.000 description 2
- 244000061456 Solanum tuberosum Species 0.000 description 2
- 235000011684 Sorghum saccharatum Nutrition 0.000 description 2
- 238000000692 Student's t-test Methods 0.000 description 2
- 235000021536 Sugar beet Nutrition 0.000 description 2
- 241000282887 Suidae Species 0.000 description 2
- 101001099217 Thermotoga maritima (strain ATCC 43589 / DSM 3109 / JCM 10099 / NBRC 100826 / MSB8) Triosephosphate isomerase Proteins 0.000 description 2
- 101710183280 Topoisomerase Proteins 0.000 description 2
- 108700009124 Transcription Initiation Site Proteins 0.000 description 2
- 208000036142 Viral infection Diseases 0.000 description 2
- 241000520892 Xanthomonas axonopodis Species 0.000 description 2
- 241000589636 Xanthomonas campestris Species 0.000 description 2
- 235000016383 Zea mays subsp huehuetenangensis Nutrition 0.000 description 2
- 230000002159 abnormal effect Effects 0.000 description 2
- 230000002378 acidificating effect Effects 0.000 description 2
- 239000002870 angiogenesis inducing agent Substances 0.000 description 2
- 239000000427 antigen Substances 0.000 description 2
- 108091007433 antigens Proteins 0.000 description 2
- 102000036639 antigens Human genes 0.000 description 2
- 230000006907 apoptotic process Effects 0.000 description 2
- 244000052616 bacterial pathogen Species 0.000 description 2
- 230000009286 beneficial effect Effects 0.000 description 2
- 210000004204 blood vessel Anatomy 0.000 description 2
- 210000001185 bone marrow Anatomy 0.000 description 2
- 239000000872 buffer Substances 0.000 description 2
- 239000011575 calcium Substances 0.000 description 2
- 239000001506 calcium phosphate Substances 0.000 description 2
- 229910000389 calcium phosphate Inorganic materials 0.000 description 2
- 235000011010 calcium phosphates Nutrition 0.000 description 2
- 125000003178 carboxy group Chemical group [H]OC(*)=O 0.000 description 2
- 238000004113 cell culture Methods 0.000 description 2
- 235000013339 cereals Nutrition 0.000 description 2
- 235000013330 chicken meat Nutrition 0.000 description 2
- 210000003763 chloroplast Anatomy 0.000 description 2
- 230000000295 complement effect Effects 0.000 description 2
- 230000021615 conjugation Effects 0.000 description 2
- 238000012937 correction Methods 0.000 description 2
- 238000005520 cutting process Methods 0.000 description 2
- 210000000805 cytoplasm Anatomy 0.000 description 2
- 238000001514 detection method Methods 0.000 description 2
- 230000004069 differentiation Effects 0.000 description 2
- 208000035475 disorder Diseases 0.000 description 2
- 229940079593 drug Drugs 0.000 description 2
- 238000009510 drug design Methods 0.000 description 2
- 235000013399 edible fruits Nutrition 0.000 description 2
- 210000002308 embryonic cell Anatomy 0.000 description 2
- 239000003623 enhancer Substances 0.000 description 2
- 230000007613 environmental effect Effects 0.000 description 2
- 230000001973 epigenetic effect Effects 0.000 description 2
- 230000001747 exhibiting effect Effects 0.000 description 2
- 235000013305 food Nutrition 0.000 description 2
- 230000004345 fruit ripening Effects 0.000 description 2
- 238000001476 gene delivery Methods 0.000 description 2
- 150000004676 glycans Chemical class 0.000 description 2
- 210000002149 gonad Anatomy 0.000 description 2
- 229910001385 heavy metal Inorganic materials 0.000 description 2
- IPCSVZSSVZVIGE-UHFFFAOYSA-N hexadecanoic acid Chemical compound CCCCCCCCCCCCCCCC(O)=O IPCSVZSSVZVIGE-UHFFFAOYSA-N 0.000 description 2
- 229940088597 hormone Drugs 0.000 description 2
- 239000005556 hormone Substances 0.000 description 2
- 210000005260 human cell Anatomy 0.000 description 2
- 238000003018 immunoassay Methods 0.000 description 2
- 238000012405 in silico analysis Methods 0.000 description 2
- 238000005462 in vivo assay Methods 0.000 description 2
- 208000015181 infectious disease Diseases 0.000 description 2
- 238000001802 infusion Methods 0.000 description 2
- 239000012212 insulator Substances 0.000 description 2
- 238000007918 intramuscular administration Methods 0.000 description 2
- 208000028867 ischemia Diseases 0.000 description 2
- 235000020778 linoleic acid Nutrition 0.000 description 2
- OYHQOLUKZRVURQ-IXWMQOLASA-N linoleic acid Natural products CCCCC\C=C/C\C=C\CCCCCCCC(O)=O OYHQOLUKZRVURQ-IXWMQOLASA-N 0.000 description 2
- 244000144972 livestock Species 0.000 description 2
- 230000004807 localization Effects 0.000 description 2
- 229920002521 macromolecule Polymers 0.000 description 2
- 235000009973 maize Nutrition 0.000 description 2
- 230000007246 mechanism Effects 0.000 description 2
- 229930182817 methionine Natural products 0.000 description 2
- 230000011987 methylation Effects 0.000 description 2
- 238000007069 methylation reaction Methods 0.000 description 2
- 238000010369 molecular cloning Methods 0.000 description 2
- 210000003205 muscle Anatomy 0.000 description 2
- 238000002703 mutagenesis Methods 0.000 description 2
- 230000007935 neutral effect Effects 0.000 description 2
- 229910052757 nitrogen Inorganic materials 0.000 description 2
- 210000000287 oocyte Anatomy 0.000 description 2
- 210000003463 organelle Anatomy 0.000 description 2
- 244000045947 parasite Species 0.000 description 2
- 125000002467 phosphate group Chemical group [H]OP(=O)(O[H])O[*] 0.000 description 2
- 230000026731 phosphorylation Effects 0.000 description 2
- 238000006366 phosphorylation reaction Methods 0.000 description 2
- 244000000003 plant pathogen Species 0.000 description 2
- 230000008488 polyadenylation Effects 0.000 description 2
- 229920001223 polyethylene glycol Polymers 0.000 description 2
- 229920001282 polysaccharide Polymers 0.000 description 2
- 239000005017 polysaccharide Substances 0.000 description 2
- 239000013641 positive control Substances 0.000 description 2
- 230000003389 potentiating effect Effects 0.000 description 2
- 102000003998 progesterone receptors Human genes 0.000 description 2
- 108090000468 progesterone receptors Proteins 0.000 description 2
- 238000000159 protein binding assay Methods 0.000 description 2
- 102000021127 protein binding proteins Human genes 0.000 description 2
- 108091011138 protein binding proteins Proteins 0.000 description 2
- 210000001938 protoplast Anatomy 0.000 description 2
- 235000015136 pumpkin Nutrition 0.000 description 2
- 238000000746 purification Methods 0.000 description 2
- 230000002285 radioactive effect Effects 0.000 description 2
- 238000007634 remodeling Methods 0.000 description 2
- 230000008263 repair mechanism Effects 0.000 description 2
- 210000001995 reticulocyte Anatomy 0.000 description 2
- 206010039073 rheumatoid arthritis Diseases 0.000 description 2
- 108091092562 ribozyme Proteins 0.000 description 2
- 235000009566 rice Nutrition 0.000 description 2
- 230000035945 sensitivity Effects 0.000 description 2
- 230000003584 silencer Effects 0.000 description 2
- 230000005783 single-strand break Effects 0.000 description 2
- 210000003491 skin Anatomy 0.000 description 2
- 102000028561 sterol response element binding proteins Human genes 0.000 description 2
- 108091009326 sterol response element binding proteins Proteins 0.000 description 2
- 238000007920 subcutaneous administration Methods 0.000 description 2
- 235000000346 sugar Nutrition 0.000 description 2
- 239000000725 suspension Substances 0.000 description 2
- 230000009897 systematic effect Effects 0.000 description 2
- 229940124597 therapeutic agent Drugs 0.000 description 2
- RYYWUUFWQRZTIU-UHFFFAOYSA-K thiophosphate Chemical group [O-]P([O-])([O-])=S RYYWUUFWQRZTIU-UHFFFAOYSA-K 0.000 description 2
- 231100000331 toxic Toxicity 0.000 description 2
- 230000002588 toxic effect Effects 0.000 description 2
- 230000037426 transcriptional repression Effects 0.000 description 2
- 230000010474 transient expression Effects 0.000 description 2
- QORWJWZARLRLPR-UHFFFAOYSA-H tricalcium bis(phosphate) Chemical compound [Ca+2].[Ca+2].[Ca+2].[O-]P([O-])([O-])=O.[O-]P([O-])([O-])=O QORWJWZARLRLPR-UHFFFAOYSA-H 0.000 description 2
- 241001529453 unidentified herpesvirus Species 0.000 description 2
- 239000004474 valine Substances 0.000 description 2
- 208000019553 vascular disease Diseases 0.000 description 2
- 239000008158 vegetable oil Substances 0.000 description 2
- 230000009385 viral infection Effects 0.000 description 2
- XLYOFNOQVPJJNP-UHFFFAOYSA-N water Substances O XLYOFNOQVPJJNP-UHFFFAOYSA-N 0.000 description 2
- OPCHFPHZPIURNA-MFERNQICSA-N (2s)-2,5-bis(3-aminopropylamino)-n-[2-(dioctadecylamino)acetyl]pentanamide Chemical compound CCCCCCCCCCCCCCCCCCN(CC(=O)NC(=O)[C@H](CCCNCCCN)NCCCN)CCCCCCCCCCCCCCCCCC OPCHFPHZPIURNA-MFERNQICSA-N 0.000 description 1
- MZOFCQQQCNRIBI-VMXHOPILSA-N (3s)-4-[[(2s)-1-[[(2s)-1-[[(1s)-1-carboxy-2-hydroxyethyl]amino]-4-methyl-1-oxopentan-2-yl]amino]-5-(diaminomethylideneamino)-1-oxopentan-2-yl]amino]-3-[[2-[[(2s)-2,6-diaminohexanoyl]amino]acetyl]amino]-4-oxobutanoic acid Chemical compound OC[C@@H](C(O)=O)NC(=O)[C@H](CC(C)C)NC(=O)[C@H](CCCN=C(N)N)NC(=O)[C@H](CC(O)=O)NC(=O)CNC(=O)[C@@H](N)CCCCN MZOFCQQQCNRIBI-VMXHOPILSA-N 0.000 description 1
- OXEDXHIBHVMDST-UHFFFAOYSA-N 12Z-octadecenoic acid Natural products CCCCCC=CCCCCCCCCCCC(O)=O OXEDXHIBHVMDST-UHFFFAOYSA-N 0.000 description 1
- JTTIOYHBNXDJOD-UHFFFAOYSA-N 2,4,6-triaminopyrimidine Chemical compound NC1=CC(N)=NC(N)=N1 JTTIOYHBNXDJOD-UHFFFAOYSA-N 0.000 description 1
- 108010020183 3-phosphoshikimate 1-carboxyvinyltransferase Proteins 0.000 description 1
- 230000005730 ADP ribosylation Effects 0.000 description 1
- 102000000452 Acetyl-CoA carboxylase Human genes 0.000 description 1
- 108010016219 Acetyl-CoA carboxylase Proteins 0.000 description 1
- 101710159080 Aconitate hydratase A Proteins 0.000 description 1
- 101710159078 Aconitate hydratase B Proteins 0.000 description 1
- 102100034544 Acyl-CoA 6-desaturase Human genes 0.000 description 1
- 229920001817 Agar Polymers 0.000 description 1
- 244000291564 Allium cepa Species 0.000 description 1
- 235000002732 Allium cepa var. cepa Nutrition 0.000 description 1
- 102100021266 Alpha-(1,6)-fucosyltransferase Human genes 0.000 description 1
- 208000024827 Alzheimer disease Diseases 0.000 description 1
- 244000099147 Ananas comosus Species 0.000 description 1
- 235000007119 Ananas comosus Nutrition 0.000 description 1
- 108020005544 Antisense RNA Proteins 0.000 description 1
- 240000007087 Apium graveolens Species 0.000 description 1
- 235000015849 Apium graveolens Dulce Group Nutrition 0.000 description 1
- 235000010591 Appio Nutrition 0.000 description 1
- 241000219194 Arabidopsis Species 0.000 description 1
- 235000017060 Arachis glabrata Nutrition 0.000 description 1
- 244000105624 Arachis hypogaea Species 0.000 description 1
- 235000010777 Arachis hypogaea Nutrition 0.000 description 1
- 235000018262 Arachis monticola Nutrition 0.000 description 1
- 241000233788 Arecaceae Species 0.000 description 1
- DCXYFEDJOCDNAF-UHFFFAOYSA-N Asparagine Natural products OC(=O)C(N)CC(N)=O DCXYFEDJOCDNAF-UHFFFAOYSA-N 0.000 description 1
- 235000005340 Asparagus officinalis Nutrition 0.000 description 1
- 241000228212 Aspergillus Species 0.000 description 1
- BHELIUBJHYAEDK-OAIUPTLZSA-N Aspoxicillin Chemical compound C1([C@H](C(=O)N[C@@H]2C(N3[C@H](C(C)(C)S[C@@H]32)C(O)=O)=O)NC(=O)[C@H](N)CC(=O)NC)=CC=C(O)C=C1 BHELIUBJHYAEDK-OAIUPTLZSA-N 0.000 description 1
- 235000007319 Avena orientalis Nutrition 0.000 description 1
- 244000075850 Avena orientalis Species 0.000 description 1
- 241000206761 Bacillariophyta Species 0.000 description 1
- 241000193738 Bacillus anthracis Species 0.000 description 1
- 208000035143 Bacterial infection Diseases 0.000 description 1
- 235000017166 Bambusa arundinacea Nutrition 0.000 description 1
- 235000017491 Bambusa tulda Nutrition 0.000 description 1
- DWRXFEITVBNRMK-UHFFFAOYSA-N Beta-D-1-Arabinofuranosylthymine Natural products O=C1NC(=O)C(C)=CN1C1C(O)C(O)C(CO)O1 DWRXFEITVBNRMK-UHFFFAOYSA-N 0.000 description 1
- 108010018763 Biotin carboxylase Proteins 0.000 description 1
- 241001536324 Botryococcus Species 0.000 description 1
- 208000003508 Botulism Diseases 0.000 description 1
- 241000219198 Brassica Species 0.000 description 1
- 235000011331 Brassica Nutrition 0.000 description 1
- 235000014698 Brassica juncea var multisecta Nutrition 0.000 description 1
- 235000006008 Brassica napus var napus Nutrition 0.000 description 1
- 235000006618 Brassica rapa subsp oleifera Nutrition 0.000 description 1
- 244000188595 Brassica sinapistrum Species 0.000 description 1
- 206010006187 Breast cancer Diseases 0.000 description 1
- 208000026310 Breast neoplasm Diseases 0.000 description 1
- 102100031650 C-X-C chemokine receptor type 4 Human genes 0.000 description 1
- 239000002126 C01EB10 - Adenosine Substances 0.000 description 1
- 101150061009 C1 gene Proteins 0.000 description 1
- OYPRJOBELJOOCE-UHFFFAOYSA-N Calcium Chemical compound [Ca] OYPRJOBELJOOCE-UHFFFAOYSA-N 0.000 description 1
- 241000222120 Candida <Saccharomycetales> Species 0.000 description 1
- 241000283707 Capra Species 0.000 description 1
- 108090000565 Capsid Proteins Proteins 0.000 description 1
- 206010007269 Carcinogenicity Diseases 0.000 description 1
- 102100028892 Cardiotrophin-1 Human genes 0.000 description 1
- 102000004018 Caspase 6 Human genes 0.000 description 1
- 108090000425 Caspase 6 Proteins 0.000 description 1
- 102000011727 Caspases Human genes 0.000 description 1
- 108010076667 Caspases Proteins 0.000 description 1
- 102000020313 Cell-Penetrating Peptides Human genes 0.000 description 1
- 108010051109 Cell-Penetrating Peptides Proteins 0.000 description 1
- 102100025064 Cellular tumor antigen p53 Human genes 0.000 description 1
- 241000282693 Cercopithecidae Species 0.000 description 1
- 241000606161 Chlamydia Species 0.000 description 1
- 241000195649 Chlorella <Chlorellales> Species 0.000 description 1
- 235000007516 Chrysanthemum Nutrition 0.000 description 1
- 244000189548 Chrysanthemum x morifolium Species 0.000 description 1
- 244000241235 Citrullus lanatus Species 0.000 description 1
- 235000012828 Citrullus lanatus var citroides Nutrition 0.000 description 1
- 241000207199 Citrus Species 0.000 description 1
- 102100022641 Coagulation factor IX Human genes 0.000 description 1
- 235000013162 Cocos nucifera Nutrition 0.000 description 1
- 244000060011 Cocos nucifera Species 0.000 description 1
- 206010009944 Colon cancer Diseases 0.000 description 1
- 241000723607 Comovirus Species 0.000 description 1
- 108020004635 Complementary DNA Proteins 0.000 description 1
- 241000218631 Coniferophyta Species 0.000 description 1
- 240000000491 Corchorus aestuans Species 0.000 description 1
- 235000011777 Corchorus aestuans Nutrition 0.000 description 1
- 235000010862 Corchorus capsularis Nutrition 0.000 description 1
- 241000709687 Coxsackievirus Species 0.000 description 1
- 241000699802 Cricetulus griseus Species 0.000 description 1
- 244000241257 Cucumis melo Species 0.000 description 1
- 235000015510 Cucumis melo subsp melo Nutrition 0.000 description 1
- 241000219122 Cucurbita Species 0.000 description 1
- 240000001980 Cucurbita pepo Species 0.000 description 1
- 241000192700 Cyanobacteria Species 0.000 description 1
- 206010011732 Cyst Diseases 0.000 description 1
- 201000003883 Cystic fibrosis Diseases 0.000 description 1
- 101710155335 DELLA protein SLR1 Proteins 0.000 description 1
- 108010009540 DNA (Cytosine-5-)-Methyltransferase 1 Proteins 0.000 description 1
- 102100024812 DNA (cytosine-5)-methyltransferase 3A Human genes 0.000 description 1
- 102100024810 DNA (cytosine-5)-methyltransferase 3B Human genes 0.000 description 1
- 101710123222 DNA (cytosine-5)-methyltransferase 3B Proteins 0.000 description 1
- 108010024491 DNA Methyltransferase 3A Proteins 0.000 description 1
- 102000011724 DNA Repair Enzymes Human genes 0.000 description 1
- 108010076525 DNA Repair Enzymes Proteins 0.000 description 1
- 230000004544 DNA amplification Effects 0.000 description 1
- 238000007400 DNA extraction Methods 0.000 description 1
- 230000008836 DNA modification Effects 0.000 description 1
- 238000001712 DNA sequencing Methods 0.000 description 1
- 108090000626 DNA-directed RNA polymerases Proteins 0.000 description 1
- 102000004163 DNA-directed RNA polymerases Human genes 0.000 description 1
- 241000725619 Dengue virus Species 0.000 description 1
- 108010054576 Deoxyribonuclease EcoRI Proteins 0.000 description 1
- 229920002307 Dextran Polymers 0.000 description 1
- 206010012689 Diabetic retinopathy Diseases 0.000 description 1
- 102100024746 Dihydrofolate reductase Human genes 0.000 description 1
- 241000255581 Drosophila <fruit fly, genus> Species 0.000 description 1
- KCXVZYZYPLLWCC-UHFFFAOYSA-N EDTA Chemical compound OC(=O)CN(CC(O)=O)CCN(CC(O)=O)CC(O)=O KCXVZYZYPLLWCC-UHFFFAOYSA-N 0.000 description 1
- 201000011001 Ebola Hemorrhagic Fever Diseases 0.000 description 1
- UPEZCKBFRMILAV-JNEQICEOSA-N Ecdysone Natural products O=C1[C@H]2[C@@](C)([C@@H]3C([C@@]4(O)[C@@](C)([C@H]([C@H]([C@@H](O)CCC(O)(C)C)C)CC4)CC3)=C1)C[C@H](O)[C@H](O)C2 UPEZCKBFRMILAV-JNEQICEOSA-N 0.000 description 1
- 241001466953 Echovirus Species 0.000 description 1
- 235000001950 Elaeis guineensis Nutrition 0.000 description 1
- 244000127993 Elaeis melanococca Species 0.000 description 1
- 101100049549 Enterobacteria phage P4 sid gene Proteins 0.000 description 1
- 101000889905 Enterobacteria phage RB3 Intron-associated endonuclease 3 Proteins 0.000 description 1
- 101000889904 Enterobacteria phage T4 Defective intron-associated endonuclease 3 Proteins 0.000 description 1
- 101000889900 Enterobacteria phage T4 Intron-associated endonuclease 1 Proteins 0.000 description 1
- 101000889899 Enterobacteria phage T4 Intron-associated endonuclease 2 Proteins 0.000 description 1
- 241000709661 Enterovirus Species 0.000 description 1
- 241000991587 Enterovirus C Species 0.000 description 1
- 101800001467 Envelope glycoprotein E2 Proteins 0.000 description 1
- 101710091045 Envelope protein Proteins 0.000 description 1
- 241001481760 Erethizon dorsatum Species 0.000 description 1
- 108060002716 Exonuclease Proteins 0.000 description 1
- 108010076282 Factor IX Proteins 0.000 description 1
- 102100034543 Fatty acid desaturase 3 Human genes 0.000 description 1
- 108010087894 Fatty acid desaturases Proteins 0.000 description 1
- 241000724791 Filamentous phage Species 0.000 description 1
- 241000710831 Flavivirus Species 0.000 description 1
- 240000009088 Fragaria x ananassa Species 0.000 description 1
- 108091092584 GDNA Proteins 0.000 description 1
- 108700007698 Genetic Terminator Regions Proteins 0.000 description 1
- 241000224466 Giardia Species 0.000 description 1
- 241000713813 Gibbon ape leukemia virus Species 0.000 description 1
- 108700023224 Glucose-1-phosphate adenylyltransferases Proteins 0.000 description 1
- DHMQDGOQFOQNFH-UHFFFAOYSA-N Glycine Chemical compound NCC(O)=O DHMQDGOQFOQNFH-UHFFFAOYSA-N 0.000 description 1
- 239000004471 Glycine Substances 0.000 description 1
- 101100446350 Glycine max FAD2-2 gene Proteins 0.000 description 1
- 102000003886 Glycoproteins Human genes 0.000 description 1
- 108090000288 Glycoproteins Proteins 0.000 description 1
- 241000206581 Gracilaria Species 0.000 description 1
- 108010017213 Granulocyte-Macrophage Colony-Stimulating Factor Proteins 0.000 description 1
- 102000004457 Granulocyte-Macrophage Colony-Stimulating Factor Human genes 0.000 description 1
- 108010043121 Green Fluorescent Proteins Proteins 0.000 description 1
- 102000004144 Green Fluorescent Proteins Human genes 0.000 description 1
- 208000031886 HIV Infections Diseases 0.000 description 1
- 108010002459 HIV Integrase Proteins 0.000 description 1
- 208000037357 HIV infectious disease Diseases 0.000 description 1
- 101710154606 Hemagglutinin Proteins 0.000 description 1
- 108010049606 Hepatocyte Nuclear Factors Proteins 0.000 description 1
- 102000008088 Hepatocyte Nuclear Factors Human genes 0.000 description 1
- 102100039869 Histone H2B type F-S Human genes 0.000 description 1
- 102100022846 Histone acetyltransferase KAT2B Human genes 0.000 description 1
- 108090000246 Histone acetyltransferases Proteins 0.000 description 1
- 102000003893 Histone acetyltransferases Human genes 0.000 description 1
- 101000819490 Homo sapiens Alpha-(1,6)-fucosyltransferase Proteins 0.000 description 1
- 101000922348 Homo sapiens C-X-C chemokine receptor type 4 Proteins 0.000 description 1
- 101100438883 Homo sapiens CCR5 gene Proteins 0.000 description 1
- 101000916283 Homo sapiens Cardiotrophin-1 Proteins 0.000 description 1
- 101000721661 Homo sapiens Cellular tumor antigen p53 Proteins 0.000 description 1
- 101000931098 Homo sapiens DNA (cytosine-5)-methyltransferase 1 Proteins 0.000 description 1
- 101000851181 Homo sapiens Epidermal growth factor receptor Proteins 0.000 description 1
- 101001035372 Homo sapiens Histone H2B type F-S Proteins 0.000 description 1
- 101001047006 Homo sapiens Histone acetyltransferase KAT2B Proteins 0.000 description 1
- 101000615495 Homo sapiens Methyl-CpG-binding domain protein 3 Proteins 0.000 description 1
- 101000724418 Homo sapiens Neutral amino acid transporter B(0) Proteins 0.000 description 1
- 101100137155 Homo sapiens POU5F1 gene Proteins 0.000 description 1
- 101001109800 Homo sapiens Pro-neuregulin-1, membrane-bound isoform Proteins 0.000 description 1
- 101000702560 Homo sapiens Probable global transcription activator SNF2L1 Proteins 0.000 description 1
- 101000738771 Homo sapiens Receptor-type tyrosine-protein phosphatase C Proteins 0.000 description 1
- 101000702544 Homo sapiens SWI/SNF-related matrix-associated actin-dependent regulator of chromatin subfamily A member 5 Proteins 0.000 description 1
- 101000802101 Homo sapiens mRNA decay activator protein ZFP36L2 Proteins 0.000 description 1
- 241000598436 Human T-cell lymphotropic virus Species 0.000 description 1
- 241000700588 Human alphaherpesvirus 1 Species 0.000 description 1
- 208000023105 Huntington disease Diseases 0.000 description 1
- 208000035150 Hypercholesterolemia Diseases 0.000 description 1
- 108010091358 Hypoxanthine Phosphoribosyltransferase Proteins 0.000 description 1
- 102100029098 Hypoxanthine-guanine phosphoribosyltransferase Human genes 0.000 description 1
- 102000044753 ISWI Human genes 0.000 description 1
- 102000008394 Immunoglobulin Fragments Human genes 0.000 description 1
- 108010021625 Immunoglobulin Fragments Proteins 0.000 description 1
- 208000026350 Inborn Genetic disease Diseases 0.000 description 1
- 102000012330 Integrases Human genes 0.000 description 1
- 102100037850 Interferon gamma Human genes 0.000 description 1
- 108010074328 Interferon-gamma Proteins 0.000 description 1
- 108091092195 Intron Proteins 0.000 description 1
- 241000588748 Klebsiella Species 0.000 description 1
- 125000003412 L-alanyl group Chemical group [H]N([H])[C@@](C([H])([H])[H])(C(=O)[*])[H] 0.000 description 1
- 108010001831 LDL receptors Proteins 0.000 description 1
- 108091026898 Leader sequence (mRNA) Proteins 0.000 description 1
- 241000589248 Legionella Species 0.000 description 1
- 208000007764 Legionnaires' Disease Diseases 0.000 description 1
- 241000222722 Leishmania <genus> Species 0.000 description 1
- 206010024238 Leptospirosis Diseases 0.000 description 1
- 108010037138 Linoleoyl-CoA Desaturase Proteins 0.000 description 1
- 235000004431 Linum usitatissimum Nutrition 0.000 description 1
- 240000006240 Linum usitatissimum Species 0.000 description 1
- 108090001030 Lipoproteins Proteins 0.000 description 1
- 102000004895 Lipoproteins Human genes 0.000 description 1
- 102100024640 Low-density lipoprotein receptor Human genes 0.000 description 1
- 108090000856 Lyases Proteins 0.000 description 1
- 102000004317 Lyases Human genes 0.000 description 1
- 208000016604 Lyme disease Diseases 0.000 description 1
- 241000218922 Magnoliophyta Species 0.000 description 1
- 244000081841 Malus domestica Species 0.000 description 1
- 235000011430 Malus pumila Nutrition 0.000 description 1
- 244000070406 Malus silvestris Species 0.000 description 1
- 235000015103 Malus silvestris Nutrition 0.000 description 1
- 240000003183 Manihot esculenta Species 0.000 description 1
- 235000016735 Manihot esculenta subsp esculenta Nutrition 0.000 description 1
- 241000712079 Measles morbillivirus Species 0.000 description 1
- 102000006890 Methyl-CpG-Binding Protein 2 Human genes 0.000 description 1
- 108010072388 Methyl-CpG-Binding Protein 2 Proteins 0.000 description 1
- 102100021291 Methyl-CpG-binding domain protein 3 Human genes 0.000 description 1
- 102000016397 Methyltransferase Human genes 0.000 description 1
- 101150076359 Mhc gene Proteins 0.000 description 1
- 241000192041 Micrococcus Species 0.000 description 1
- 241000713869 Moloney murine leukemia virus Species 0.000 description 1
- 241000711386 Mumps virus Species 0.000 description 1
- 108010086093 Mung Bean Nuclease Proteins 0.000 description 1
- 241000699660 Mus musculus Species 0.000 description 1
- 240000005561 Musa balbisiana Species 0.000 description 1
- 235000018290 Musa x paradisiaca Nutrition 0.000 description 1
- 125000001429 N-terminal alpha-amino-acid group Chemical group 0.000 description 1
- 125000000729 N-terminal amino-acid group Chemical group 0.000 description 1
- 241000588652 Neisseria gonorrhoeae Species 0.000 description 1
- 241000244206 Nematoda Species 0.000 description 1
- 206010061309 Neoplasm progression Diseases 0.000 description 1
- 208000009869 Neu-Laxova syndrome Diseases 0.000 description 1
- 102100028267 Neutral amino acid transporter B(0) Human genes 0.000 description 1
- 235000002637 Nicotiana tabacum Nutrition 0.000 description 1
- 244000061176 Nicotiana tabacum Species 0.000 description 1
- 108010008964 Non-Histone Chromosomal Proteins Proteins 0.000 description 1
- 102000006570 Non-Histone Chromosomal Proteins Human genes 0.000 description 1
- 108091092724 Noncoding DNA Proteins 0.000 description 1
- 108020004485 Nonsense Codon Proteins 0.000 description 1
- 108700020796 Oncogene Proteins 0.000 description 1
- 108700026244 Open Reading Frames Proteins 0.000 description 1
- 101100046877 Oryza sativa subsp. japonica TRAB1 gene Proteins 0.000 description 1
- 101710093908 Outer capsid protein VP4 Proteins 0.000 description 1
- 101710135467 Outer capsid protein sigma-1 Proteins 0.000 description 1
- 238000010222 PCR analysis Methods 0.000 description 1
- 229910019142 PO4 Inorganic materials 0.000 description 1
- 102100035593 POU domain, class 2, transcription factor 1 Human genes 0.000 description 1
- 101710084414 POU domain, class 2, transcription factor 1 Proteins 0.000 description 1
- 235000021314 Palmitic acid Nutrition 0.000 description 1
- 241000282376 Panthera tigris Species 0.000 description 1
- 241001631646 Papillomaviridae Species 0.000 description 1
- 239000006002 Pepper Substances 0.000 description 1
- 101100440941 Petroselinum crispum CPRF1 gene Proteins 0.000 description 1
- 240000007377 Petunia x hybrida Species 0.000 description 1
- 235000010627 Phaseolus vulgaris Nutrition 0.000 description 1
- 244000046052 Phaseolus vulgaris Species 0.000 description 1
- 235000015334 Phyllostachys viridis Nutrition 0.000 description 1
- 244000082204 Phyllostachys viridis Species 0.000 description 1
- 241000218657 Picea Species 0.000 description 1
- 241000235648 Pichia Species 0.000 description 1
- 235000016761 Piper aduncum Nutrition 0.000 description 1
- 235000017804 Piper guineense Nutrition 0.000 description 1
- 240000003889 Piper guineense Species 0.000 description 1
- 235000008184 Piper nigrum Nutrition 0.000 description 1
- 241000758706 Piperaceae Species 0.000 description 1
- 235000010582 Pisum sativum Nutrition 0.000 description 1
- 240000004713 Pisum sativum Species 0.000 description 1
- 206010035148 Plague Diseases 0.000 description 1
- 101000658568 Planomicrobium okeanokoites Type II restriction enzyme FokI Proteins 0.000 description 1
- 206010035226 Plasma cell myeloma Diseases 0.000 description 1
- 108010059820 Polygalacturonase Proteins 0.000 description 1
- 241000219000 Populus Species 0.000 description 1
- 241000288906 Primates Species 0.000 description 1
- 101710176177 Protein A56 Proteins 0.000 description 1
- 101710188315 Protein X Proteins 0.000 description 1
- 241000588769 Proteus <enterobacteria> Species 0.000 description 1
- 206010037075 Protozoal infections Diseases 0.000 description 1
- 241000589516 Pseudomonas Species 0.000 description 1
- 201000004681 Psoriasis Diseases 0.000 description 1
- 241000220324 Pyrus Species 0.000 description 1
- 102000044126 RNA-Binding Proteins Human genes 0.000 description 1
- 230000004570 RNA-binding Effects 0.000 description 1
- 101710105008 RNA-binding protein Proteins 0.000 description 1
- 101150065817 ROM2 gene Proteins 0.000 description 1
- 241000711798 Rabies lyssavirus Species 0.000 description 1
- 101100272715 Ralstonia solanacearum (strain GMI1000) brg11 gene Proteins 0.000 description 1
- 235000005733 Raphanus sativus var niger Nutrition 0.000 description 1
- 244000155437 Raphanus sativus var. niger Species 0.000 description 1
- 101000702488 Rattus norvegicus High affinity cationic amino acid transporter 1 Proteins 0.000 description 1
- 102100037422 Receptor-type tyrosine-protein phosphatase C Human genes 0.000 description 1
- 108020004511 Recombinant DNA Proteins 0.000 description 1
- 241000242739 Renilla Species 0.000 description 1
- 241000725643 Respiratory syncytial virus Species 0.000 description 1
- 241000589187 Rhizobium sp. Species 0.000 description 1
- 241000606701 Rickettsia Species 0.000 description 1
- 241000283984 Rodentia Species 0.000 description 1
- 235000004789 Rosa xanthina Nutrition 0.000 description 1
- 241000109329 Rosa xanthina Species 0.000 description 1
- 241000702670 Rotavirus Species 0.000 description 1
- 241000710799 Rubella virus Species 0.000 description 1
- 238000011579 SCID mouse model Methods 0.000 description 1
- 241000235070 Saccharomyces Species 0.000 description 1
- 101001025539 Saccharomyces cerevisiae (strain ATCC 204508 / S288c) Homothallic switching endonuclease Proteins 0.000 description 1
- 240000000111 Saccharum officinarum Species 0.000 description 1
- 235000007201 Saccharum officinarum Nutrition 0.000 description 1
- 241000717635 Salix sitchensis Species 0.000 description 1
- 241000607142 Salmonella Species 0.000 description 1
- 241000235346 Schizosaccharomyces Species 0.000 description 1
- 241000245026 Scoliopus bigelovii Species 0.000 description 1
- 241000209056 Secale Species 0.000 description 1
- 244000082988 Secale cereale Species 0.000 description 1
- MTCFGRXMJLQNBG-UHFFFAOYSA-N Serine Natural products OCC(N)C(O)=O MTCFGRXMJLQNBG-UHFFFAOYSA-N 0.000 description 1
- 241000607720 Serratia Species 0.000 description 1
- XUIMIQQOPSSXEZ-UHFFFAOYSA-N Silicon Chemical compound [Si] XUIMIQQOPSSXEZ-UHFFFAOYSA-N 0.000 description 1
- 241000700584 Simplexvirus Species 0.000 description 1
- 235000009337 Spinacia oleracea Nutrition 0.000 description 1
- 244000300264 Spinacia oleracea Species 0.000 description 1
- 241000256251 Spodoptera frugiperda Species 0.000 description 1
- 241000295644 Staphylococcaceae Species 0.000 description 1
- 108010039811 Starch synthase Proteins 0.000 description 1
- 102000007451 Steroid Receptors Human genes 0.000 description 1
- 108010085012 Steroid Receptors Proteins 0.000 description 1
- 208000006011 Stroke Diseases 0.000 description 1
- 108010043934 Sucrose synthase Proteins 0.000 description 1
- 101800001271 Surface protein Proteins 0.000 description 1
- 241000282898 Sus scrofa Species 0.000 description 1
- 206010043376 Tetanus Diseases 0.000 description 1
- 108091036066 Three prime untranslated region Proteins 0.000 description 1
- 108010073062 Transcription Activator-Like Effectors Proteins 0.000 description 1
- 241000224526 Trichomonas Species 0.000 description 1
- 239000007983 Tris buffer Substances 0.000 description 1
- 241000223104 Trypanosoma Species 0.000 description 1
- 108060008682 Tumor Necrosis Factor Proteins 0.000 description 1
- 102000000852 Tumor Necrosis Factor-alpha Human genes 0.000 description 1
- 108700025716 Tumor Suppressor Genes Proteins 0.000 description 1
- 102000044209 Tumor Suppressor Genes Human genes 0.000 description 1
- 241000700618 Vaccinia virus Species 0.000 description 1
- 241000607626 Vibrio cholerae Species 0.000 description 1
- 241000219977 Vigna Species 0.000 description 1
- 235000010726 Vigna sinensis Nutrition 0.000 description 1
- 108020005202 Viral DNA Proteins 0.000 description 1
- 235000009754 Vitis X bourquina Nutrition 0.000 description 1
- 235000012333 Vitis X labruscana Nutrition 0.000 description 1
- 235000014787 Vitis vinifera Nutrition 0.000 description 1
- 240000006365 Vitis vinifera Species 0.000 description 1
- 241000607479 Yersinia pestis Species 0.000 description 1
- PTFCDOFLOPIGGS-UHFFFAOYSA-N Zinc dication Chemical compound [Zn+2] PTFCDOFLOPIGGS-UHFFFAOYSA-N 0.000 description 1
- HMNZFMSWFCAGGW-XPWSMXQVSA-N [3-[hydroxy(2-hydroxyethoxy)phosphoryl]oxy-2-[(e)-octadec-9-enoyl]oxypropyl] (e)-octadec-9-enoate Chemical compound CCCCCCCC\C=C\CCCCCCCC(=O)OCC(COP(O)(=O)OCCO)OC(=O)CCCCCCC\C=C\CCCCCCCC HMNZFMSWFCAGGW-XPWSMXQVSA-N 0.000 description 1
- FJJCIZWZNKZHII-UHFFFAOYSA-N [4,6-bis(cyanoamino)-1,3,5-triazin-2-yl]cyanamide Chemical compound N#CNC1=NC(NC#N)=NC(NC#N)=N1 FJJCIZWZNKZHII-UHFFFAOYSA-N 0.000 description 1
- 230000021736 acetylation Effects 0.000 description 1
- 238000006640 acetylation reaction Methods 0.000 description 1
- 108020002494 acetyltransferase Proteins 0.000 description 1
- 102000005421 acetyltransferase Human genes 0.000 description 1
- 108700021044 acyl-ACP thioesterase Proteins 0.000 description 1
- OIRDTQYFTABQOQ-KQYNXXCUSA-N adenosine Chemical compound C1=NC=2C(N)=NC=NC=2N1[C@@H]1O[C@H](CO)[C@@H](O)[C@H]1O OIRDTQYFTABQOQ-KQYNXXCUSA-N 0.000 description 1
- 229960005305 adenosine Drugs 0.000 description 1
- 210000004504 adult stem cell Anatomy 0.000 description 1
- 239000008272 agar Substances 0.000 description 1
- 244000193174 agave Species 0.000 description 1
- 230000032683 aging Effects 0.000 description 1
- UPEZCKBFRMILAV-UHFFFAOYSA-N alpha-Ecdysone Natural products C1C(O)C(O)CC2(C)C(CCC3(C(C(C(O)CCC(C)(C)O)C)CCC33O)C)C3=CC(=O)C21 UPEZCKBFRMILAV-UHFFFAOYSA-N 0.000 description 1
- DTOSIQBPPRVQHS-PDBXOOCHSA-N alpha-linolenic acid Chemical compound CC\C=C/C\C=C/C\C=C/CCCCCCCC(O)=O DTOSIQBPPRVQHS-PDBXOOCHSA-N 0.000 description 1
- 235000020661 alpha-linolenic acid Nutrition 0.000 description 1
- 230000001772 anti-angiogenic effect Effects 0.000 description 1
- 230000000692 anti-sense effect Effects 0.000 description 1
- 230000005809 anti-tumor immunity Effects 0.000 description 1
- 210000000612 antigen-presenting cell Anatomy 0.000 description 1
- 239000003963 antioxidant agent Substances 0.000 description 1
- 235000021016 apples Nutrition 0.000 description 1
- 206010003246 arthritis Diseases 0.000 description 1
- 235000009582 asparagine Nutrition 0.000 description 1
- 235000003704 aspartic acid Nutrition 0.000 description 1
- 238000002820 assay format Methods 0.000 description 1
- VJBCNMFKFZIXHC-UHFFFAOYSA-N azanium;2-(4-methyl-5-oxo-4-propan-2-yl-1h-imidazol-2-yl)quinoline-3-carboxylate Chemical group N.N1C(=O)C(C(C)C)(C)N=C1C1=NC2=CC=CC=C2C=C1C(O)=O VJBCNMFKFZIXHC-UHFFFAOYSA-N 0.000 description 1
- 229940065181 bacillus anthracis Drugs 0.000 description 1
- 208000022362 bacterial infectious disease Diseases 0.000 description 1
- 239000000022 bacteriostatic agent Substances 0.000 description 1
- 239000011425 bamboo Substances 0.000 description 1
- 239000011324 bead Substances 0.000 description 1
- 102000005936 beta-Galactosidase Human genes 0.000 description 1
- 108010005774 beta-Galactosidase Proteins 0.000 description 1
- IQFYYKKMVGJFEH-UHFFFAOYSA-N beta-L-thymidine Natural products O=C1NC(=O)C(C)=CN1C1OC(CO)C(O)C1 IQFYYKKMVGJFEH-UHFFFAOYSA-N 0.000 description 1
- OQFSQFPPLPISGP-UHFFFAOYSA-N beta-carboxyaspartic acid Natural products OC(=O)C(N)C(C(O)=O)C(O)=O OQFSQFPPLPISGP-UHFFFAOYSA-N 0.000 description 1
- 238000002306 biochemical method Methods 0.000 description 1
- 230000003115 biocidal effect Effects 0.000 description 1
- 230000004071 biological effect Effects 0.000 description 1
- 230000031018 biological processes and functions Effects 0.000 description 1
- 230000005540 biological transmission Effects 0.000 description 1
- 229960000074 biopharmaceutical Drugs 0.000 description 1
- 238000001574 biopsy Methods 0.000 description 1
- 230000000903 blocking effect Effects 0.000 description 1
- 210000000601 blood cell Anatomy 0.000 description 1
- 210000002798 bone marrow cell Anatomy 0.000 description 1
- 210000004556 brain Anatomy 0.000 description 1
- 235000012467 brownies Nutrition 0.000 description 1
- 210000004900 c-terminal fragment Anatomy 0.000 description 1
- 229910052791 calcium Inorganic materials 0.000 description 1
- 238000004422 calculation algorithm Methods 0.000 description 1
- 150000001720 carbohydrates Chemical class 0.000 description 1
- 235000014633 carbohydrates Nutrition 0.000 description 1
- 229910052799 carbon Inorganic materials 0.000 description 1
- 230000007670 carcinogenicity Effects 0.000 description 1
- 231100000260 carcinogenicity Toxicity 0.000 description 1
- 125000002091 cationic group Chemical group 0.000 description 1
- 230000001364 causal effect Effects 0.000 description 1
- 230000032823 cell division Effects 0.000 description 1
- 230000007910 cell fusion Effects 0.000 description 1
- 230000010261 cell growth Effects 0.000 description 1
- 230000010307 cell transformation Effects 0.000 description 1
- 230000019522 cellular metabolic process Effects 0.000 description 1
- 108010040093 cellulose synthase Proteins 0.000 description 1
- 238000012512 characterization method Methods 0.000 description 1
- 239000002738 chelating agent Substances 0.000 description 1
- 230000007073 chemical hydrolysis Effects 0.000 description 1
- 235000020971 citrus fruits Nutrition 0.000 description 1
- 238000003501 co-culture Methods 0.000 description 1
- 238000000749 co-immunoprecipitation Methods 0.000 description 1
- 238000000975 co-precipitation Methods 0.000 description 1
- 230000003081 coactivator Effects 0.000 description 1
- 210000001072 colon Anatomy 0.000 description 1
- 208000029742 colonic neoplasm Diseases 0.000 description 1
- 238000007398 colorimetric assay Methods 0.000 description 1
- 239000003184 complementary RNA Substances 0.000 description 1
- 108091036078 conserved sequence Proteins 0.000 description 1
- 238000011109 contamination Methods 0.000 description 1
- 239000013068 control sample Substances 0.000 description 1
- 230000001276 controlling effect Effects 0.000 description 1
- 238000010411 cooking Methods 0.000 description 1
- 238000001816 cooling Methods 0.000 description 1
- 235000012343 cottonseed oil Nutrition 0.000 description 1
- 230000008878 coupling Effects 0.000 description 1
- 238000010168 coupling process Methods 0.000 description 1
- 238000005859 coupling reaction Methods 0.000 description 1
- 239000013078 crystal Substances 0.000 description 1
- 238000012136 culture method Methods 0.000 description 1
- 210000000448 cultured tumor cell Anatomy 0.000 description 1
- 208000031513 cyst Diseases 0.000 description 1
- 230000001086 cytosolic effect Effects 0.000 description 1
- 238000012350 deep sequencing Methods 0.000 description 1
- 238000002716 delivery method Methods 0.000 description 1
- 108010005155 delta-12 fatty acid desaturase Proteins 0.000 description 1
- 108010011713 delta-15 desaturase Proteins 0.000 description 1
- 239000005547 deoxyribonucleotide Substances 0.000 description 1
- 125000002637 deoxyribonucleotide group Chemical group 0.000 description 1
- 239000003599 detergent Substances 0.000 description 1
- 238000003745 diagnosis Methods 0.000 description 1
- 238000002405 diagnostic procedure Methods 0.000 description 1
- 229910003460 diamond Inorganic materials 0.000 description 1
- 239000010432 diamond Substances 0.000 description 1
- 108020001096 dihydrofolate reductase Proteins 0.000 description 1
- 229910001873 dinitrogen Inorganic materials 0.000 description 1
- 235000004879 dioscorea Nutrition 0.000 description 1
- 206010013023 diphtheria Diseases 0.000 description 1
- 238000007877 drug screening Methods 0.000 description 1
- UPEZCKBFRMILAV-JMZLNJERSA-N ecdysone Chemical compound C1[C@@H](O)[C@@H](O)C[C@]2(C)[C@@H](CC[C@@]3([C@@H]([C@@H]([C@H](O)CCC(C)(C)O)C)CC[C@]33O)C)C3=CC(=O)[C@@H]21 UPEZCKBFRMILAV-JMZLNJERSA-N 0.000 description 1
- 230000013020 embryo development Effects 0.000 description 1
- 206010014599 encephalitis Diseases 0.000 description 1
- 230000012202 endocytosis Effects 0.000 description 1
- 210000002889 endothelial cell Anatomy 0.000 description 1
- 230000002255 enzymatic effect Effects 0.000 description 1
- 230000007071 enzymatic hydrolysis Effects 0.000 description 1
- 238000006047 enzymatic hydrolysis reaction Methods 0.000 description 1
- 238000001952 enzyme assay Methods 0.000 description 1
- 230000008995 epigenetic change Effects 0.000 description 1
- 230000004049 epigenetic modification Effects 0.000 description 1
- 210000002304 esc Anatomy 0.000 description 1
- 239000006277 exogenous ligand Substances 0.000 description 1
- 102000013165 exonuclease Human genes 0.000 description 1
- 239000000284 extract Substances 0.000 description 1
- 229960004222 factor ix Drugs 0.000 description 1
- 239000000835 fiber Substances 0.000 description 1
- 210000002950 fibroblast Anatomy 0.000 description 1
- 239000000796 flavoring agent Substances 0.000 description 1
- 235000019634 flavors Nutrition 0.000 description 1
- 238000000684 flow cytometry Methods 0.000 description 1
- 239000012530 fluid Substances 0.000 description 1
- 239000004459 forage Substances 0.000 description 1
- 231100000221 frame shift mutation induction Toxicity 0.000 description 1
- 230000037433 frameshift Effects 0.000 description 1
- 238000007710 freezing Methods 0.000 description 1
- 230000008014 freezing Effects 0.000 description 1
- 208000024386 fungal infectious disease Diseases 0.000 description 1
- 238000001502 gel electrophoresis Methods 0.000 description 1
- 238000012246 gene addition Methods 0.000 description 1
- 238000003209 gene knockout Methods 0.000 description 1
- 102000034356 gene-regulatory proteins Human genes 0.000 description 1
- 108091006104 gene-regulatory proteins Proteins 0.000 description 1
- 238000007429 general method Methods 0.000 description 1
- 230000004077 genetic alteration Effects 0.000 description 1
- 231100000118 genetic alteration Toxicity 0.000 description 1
- 208000016361 genetic disease Diseases 0.000 description 1
- 235000013922 glutamic acid Nutrition 0.000 description 1
- 239000004220 glutamic acid Substances 0.000 description 1
- 230000013595 glycosylation Effects 0.000 description 1
- 238000006206 glycosylation reaction Methods 0.000 description 1
- PCHJSUWPFVWCPO-UHFFFAOYSA-N gold Chemical compound [Au] PCHJSUWPFVWCPO-UHFFFAOYSA-N 0.000 description 1
- 239000010931 gold Substances 0.000 description 1
- 229910052737 gold Inorganic materials 0.000 description 1
- 239000008187 granular material Substances 0.000 description 1
- 210000003714 granulocyte Anatomy 0.000 description 1
- 235000021384 green leafy vegetables Nutrition 0.000 description 1
- 230000012010 growth Effects 0.000 description 1
- 230000009036 growth inhibition Effects 0.000 description 1
- 230000036541 health Effects 0.000 description 1
- 238000010438 heat treatment Methods 0.000 description 1
- 210000002443 helper t lymphocyte Anatomy 0.000 description 1
- 239000000185 hemagglutinin Substances 0.000 description 1
- 230000011132 hemopoiesis Effects 0.000 description 1
- 210000003897 hepatic stem cell Anatomy 0.000 description 1
- 208000006454 hepatitis Diseases 0.000 description 1
- 231100000283 hepatitis Toxicity 0.000 description 1
- 230000010196 hermaphroditism Effects 0.000 description 1
- 238000005734 heterodimerization reaction Methods 0.000 description 1
- 230000006195 histone acetylation Effects 0.000 description 1
- 102000055650 human NRG1 Human genes 0.000 description 1
- 102000057714 human NTF3 Human genes 0.000 description 1
- 208000033519 human immunodeficiency virus infectious disease Diseases 0.000 description 1
- 229930195733 hydrocarbon Natural products 0.000 description 1
- 150000002430 hydrocarbons Chemical class 0.000 description 1
- 238000006460 hydrolysis reaction Methods 0.000 description 1
- 108010064894 hydroperoxide lyase Proteins 0.000 description 1
- 210000002865 immune cell Anatomy 0.000 description 1
- 230000001900 immune effect Effects 0.000 description 1
- 230000002055 immunohistochemical effect Effects 0.000 description 1
- 238000001114 immunoprecipitation Methods 0.000 description 1
- 230000006872 improvement Effects 0.000 description 1
- 238000007901 in situ hybridization Methods 0.000 description 1
- 238000010348 incorporation Methods 0.000 description 1
- 210000004263 induced pluripotent stem cell Anatomy 0.000 description 1
- 230000006698 induction Effects 0.000 description 1
- 230000002458 infectious effect Effects 0.000 description 1
- 230000002757 inflammatory effect Effects 0.000 description 1
- 230000028709 inflammatory response Effects 0.000 description 1
- 230000002401 inhibitory effect Effects 0.000 description 1
- 230000005764 inhibitory process Effects 0.000 description 1
- 238000011850 initial investigation Methods 0.000 description 1
- 230000000977 initiatory effect Effects 0.000 description 1
- 239000000138 intercalating agent Substances 0.000 description 1
- 230000003834 intracellular effect Effects 0.000 description 1
- 238000007917 intracranial administration Methods 0.000 description 1
- 238000010255 intramuscular injection Methods 0.000 description 1
- 239000007927 intramuscular injection Substances 0.000 description 1
- 238000007912 intraperitoneal administration Methods 0.000 description 1
- 238000002955 isolation Methods 0.000 description 1
- 229960000310 isoleucine Drugs 0.000 description 1
- 238000005304 joining Methods 0.000 description 1
- 210000003734 kidney Anatomy 0.000 description 1
- 210000003292 kidney cell Anatomy 0.000 description 1
- 235000021374 legumes Nutrition 0.000 description 1
- 108020001756 ligand binding domains Proteins 0.000 description 1
- 230000000670 limiting effect Effects 0.000 description 1
- 229960004488 linolenic acid Drugs 0.000 description 1
- KQQKGWQCNNTQJW-UHFFFAOYSA-N linolenic acid Natural products CC=CCCC=CCC=CCCCCCCCC(O)=O KQQKGWQCNNTQJW-UHFFFAOYSA-N 0.000 description 1
- 239000007788 liquid Substances 0.000 description 1
- 230000007774 longterm Effects 0.000 description 1
- 239000000314 lubricant Substances 0.000 description 1
- 238000004020 luminiscence type Methods 0.000 description 1
- 210000004698 lymphocyte Anatomy 0.000 description 1
- 102100034703 mRNA decay activator protein ZFP36L2 Human genes 0.000 description 1
- 208000002780 macular degeneration Diseases 0.000 description 1
- 201000004792 malaria Diseases 0.000 description 1
- 235000013310 margarine Nutrition 0.000 description 1
- 239000003264 margarine Substances 0.000 description 1
- 230000013011 mating Effects 0.000 description 1
- 238000002844 melting Methods 0.000 description 1
- 230000008018 melting Effects 0.000 description 1
- 239000012528 membrane Substances 0.000 description 1
- 230000034217 membrane fusion Effects 0.000 description 1
- MYWUZJCMWCOHBA-VIFPVBQESA-N methamphetamine Chemical compound CN[C@@H](C)CC1=CC=CC=C1 MYWUZJCMWCOHBA-VIFPVBQESA-N 0.000 description 1
- 125000002496 methyl group Chemical group [H]C([H])([H])* 0.000 description 1
- 238000002493 microarray Methods 0.000 description 1
- 244000005700 microbiome Species 0.000 description 1
- 239000011859 microparticle Substances 0.000 description 1
- 230000003228 microsomal effect Effects 0.000 description 1
- 235000019713 millet Nutrition 0.000 description 1
- 230000003278 mimic effect Effects 0.000 description 1
- 210000003470 mitochondria Anatomy 0.000 description 1
- 230000002438 mitochondrial effect Effects 0.000 description 1
- 238000012544 monitoring process Methods 0.000 description 1
- VYQNWZOUAUKGHI-UHFFFAOYSA-N monobenzone Chemical compound C1=CC(O)=CC=C1OCC1=CC=CC=C1 VYQNWZOUAUKGHI-UHFFFAOYSA-N 0.000 description 1
- 238000010172 mouse model Methods 0.000 description 1
- 210000000663 muscle cell Anatomy 0.000 description 1
- 201000006938 muscular dystrophy Diseases 0.000 description 1
- 210000000066 myeloid cell Anatomy 0.000 description 1
- 201000000050 myeloid neoplasm Diseases 0.000 description 1
- WQEPLUUGTLDZJY-UHFFFAOYSA-N n-Pentadecanoic acid Natural products CCCCCCCCCCCCCCC(O)=O WQEPLUUGTLDZJY-UHFFFAOYSA-N 0.000 description 1
- 239000013642 negative control Substances 0.000 description 1
- 230000001537 neural effect Effects 0.000 description 1
- 230000004770 neurodegeneration Effects 0.000 description 1
- 208000015122 neurodegenerative disease Diseases 0.000 description 1
- 230000007823 neuropathy Effects 0.000 description 1
- 201000001119 neuropathy Diseases 0.000 description 1
- 230000005937 nuclear translocation Effects 0.000 description 1
- 230000025308 nuclear transport Effects 0.000 description 1
- 238000007899 nucleic acid hybridization Methods 0.000 description 1
- 235000016709 nutrition Nutrition 0.000 description 1
- 210000000056 organ Anatomy 0.000 description 1
- 230000008520 organization Effects 0.000 description 1
- 210000001672 ovary Anatomy 0.000 description 1
- 230000003647 oxidation Effects 0.000 description 1
- 238000007254 oxidation reaction Methods 0.000 description 1
- 238000007911 parenteral administration Methods 0.000 description 1
- 244000052769 pathogen Species 0.000 description 1
- 230000007170 pathology Effects 0.000 description 1
- 230000035778 pathophysiological process Effects 0.000 description 1
- 235000020232 peanut Nutrition 0.000 description 1
- 235000021017 pears Nutrition 0.000 description 1
- 239000008188 pellet Substances 0.000 description 1
- 208000033808 peripheral neuropathy Diseases 0.000 description 1
- 239000000546 pharmaceutical excipient Substances 0.000 description 1
- NBIIXXVUZAFLBC-UHFFFAOYSA-K phosphate Chemical compound [O-]P([O-])([O-])=O NBIIXXVUZAFLBC-UHFFFAOYSA-K 0.000 description 1
- 239000010452 phosphate Substances 0.000 description 1
- 230000035790 physiological processes and functions Effects 0.000 description 1
- 108090000102 pigment epithelium-derived factor Proteins 0.000 description 1
- 230000008121 plant development Effects 0.000 description 1
- 230000037039 plant physiology Effects 0.000 description 1
- 239000013600 plasmid vector Substances 0.000 description 1
- 239000004033 plastic Substances 0.000 description 1
- 229920003023 plastic Polymers 0.000 description 1
- 229920001983 poloxamer Polymers 0.000 description 1
- 235000021085 polyunsaturated fats Nutrition 0.000 description 1
- 235000012015 potatoes Nutrition 0.000 description 1
- 239000000843 powder Substances 0.000 description 1
- 238000001556 precipitation Methods 0.000 description 1
- 239000002243 precursor Substances 0.000 description 1
- 239000003755 preservative agent Substances 0.000 description 1
- 210000001236 prokaryotic cell Anatomy 0.000 description 1
- 230000000069 prophylactic effect Effects 0.000 description 1
- 230000004952 protein activity Effects 0.000 description 1
- 230000012846 protein folding Effects 0.000 description 1
- 230000004853 protein function Effects 0.000 description 1
- 238000001742 protein purification Methods 0.000 description 1
- 238000011002 quantification Methods 0.000 description 1
- 238000002708 random mutagenesis Methods 0.000 description 1
- 230000009257 reactivity Effects 0.000 description 1
- 230000009467 reduction Effects 0.000 description 1
- 230000008844 regulatory mechanism Effects 0.000 description 1
- 230000003362 replicative effect Effects 0.000 description 1
- 238000007894 restriction fragment length polymorphism technique Methods 0.000 description 1
- 230000000717 retained effect Effects 0.000 description 1
- 230000002441 reversible effect Effects 0.000 description 1
- 229920002477 rna polymer Polymers 0.000 description 1
- 235000012045 salad Nutrition 0.000 description 1
- 150000003839 salts Chemical class 0.000 description 1
- 238000007480 sanger sequencing Methods 0.000 description 1
- 230000028327 secretion Effects 0.000 description 1
- 238000005204 segregation Methods 0.000 description 1
- 238000010187 selection method Methods 0.000 description 1
- 102000023888 sequence-specific DNA binding proteins Human genes 0.000 description 1
- 108091008420 sequence-specific DNA binding proteins Proteins 0.000 description 1
- 238000004904 shortening Methods 0.000 description 1
- 208000007056 sickle cell anemia Diseases 0.000 description 1
- 229910052710 silicon Inorganic materials 0.000 description 1
- 239000010703 silicon Substances 0.000 description 1
- HBMJWWWQQXIZIP-UHFFFAOYSA-N silicon carbide Chemical compound [Si+]#[C-] HBMJWWWQQXIZIP-UHFFFAOYSA-N 0.000 description 1
- 229910010271 silicon carbide Inorganic materials 0.000 description 1
- 238000002741 site-directed mutagenesis Methods 0.000 description 1
- 238000004513 sizing Methods 0.000 description 1
- 239000011780 sodium chloride Substances 0.000 description 1
- 239000002904 solvent Substances 0.000 description 1
- 210000001082 somatic cell Anatomy 0.000 description 1
- 238000010374 somatic cell nuclear transfer Methods 0.000 description 1
- 230000009870 specific binding Effects 0.000 description 1
- 238000001228 spectrum Methods 0.000 description 1
- 230000002269 spontaneous effect Effects 0.000 description 1
- 230000006641 stabilisation Effects 0.000 description 1
- 238000011105 stabilization Methods 0.000 description 1
- 239000003381 stabilizer Substances 0.000 description 1
- 230000001954 sterilising effect Effects 0.000 description 1
- 150000003431 steroids Chemical class 0.000 description 1
- 235000021012 strawberries Nutrition 0.000 description 1
- 239000006228 supernatant Substances 0.000 description 1
- 230000004083 survival effect Effects 0.000 description 1
- 239000000375 suspending agent Substances 0.000 description 1
- 208000024891 symptom Diseases 0.000 description 1
- 230000002195 synergetic effect Effects 0.000 description 1
- 238000003786 synthesis reaction Methods 0.000 description 1
- 230000002194 synthesizing effect Effects 0.000 description 1
- 239000003826 tablet Substances 0.000 description 1
- 208000001608 teratocarcinoma Diseases 0.000 description 1
- 238000002560 therapeutic procedure Methods 0.000 description 1
- 239000002562 thickening agent Substances 0.000 description 1
- 229940104230 thymidine Drugs 0.000 description 1
- 230000005026 transcription initiation Effects 0.000 description 1
- 238000011426 transformation method Methods 0.000 description 1
- 238000011830 transgenic mouse model Methods 0.000 description 1
- 230000001052 transient effect Effects 0.000 description 1
- 230000005945 translocation Effects 0.000 description 1
- 230000032258 transport Effects 0.000 description 1
- 125000003203 triacylglycerol group Chemical group 0.000 description 1
- 238000009966 trimming Methods 0.000 description 1
- LENZDBCJOHFCAS-UHFFFAOYSA-N tris Chemical compound OCC(N)(CO)CO LENZDBCJOHFCAS-UHFFFAOYSA-N 0.000 description 1
- 230000005747 tumor angiogenesis Effects 0.000 description 1
- 230000005751 tumor progression Effects 0.000 description 1
- 238000003160 two-hybrid assay Methods 0.000 description 1
- 238000010396 two-hybrid screening Methods 0.000 description 1
- 238000010798 ubiquitination Methods 0.000 description 1
- 230000034512 ubiquitination Effects 0.000 description 1
- 241000712461 unidentified influenza virus Species 0.000 description 1
- 235000013311 vegetables Nutrition 0.000 description 1
- 238000012795 verification Methods 0.000 description 1
- 210000003501 vero cell Anatomy 0.000 description 1
- 230000007998 vessel formation Effects 0.000 description 1
- 229940118696 vibrio cholerae Drugs 0.000 description 1
- 210000002845 virion Anatomy 0.000 description 1
- 239000000277 virosome Substances 0.000 description 1
- 230000001018 virulence Effects 0.000 description 1
- 238000001262 western blot Methods 0.000 description 1
Images
Classifications
-
- C—CHEMISTRY; METALLURGY
- C07—ORGANIC CHEMISTRY
- C07K—PEPTIDES
- C07K19/00—Hybrid peptides, i.e. peptides covalently bound to nucleic acids, or non-covalently bound protein-protein complexes
-
- C—CHEMISTRY; METALLURGY
- C12—BIOCHEMISTRY; BEER; SPIRITS; WINE; VINEGAR; MICROBIOLOGY; ENZYMOLOGY; MUTATION OR GENETIC ENGINEERING
- C12N—MICROORGANISMS OR ENZYMES; COMPOSITIONS THEREOF; PROPAGATING, PRESERVING, OR MAINTAINING MICROORGANISMS; MUTATION OR GENETIC ENGINEERING; CULTURE MEDIA
- C12N15/00—Mutation or genetic engineering; DNA or RNA concerning genetic engineering, vectors, e.g. plasmids, or their isolation, preparation or purification; Use of hosts therefor
- C12N15/09—Recombinant DNA-technology
- C12N15/87—Introduction of foreign genetic material using processes not otherwise provided for, e.g. co-transformation
- C12N15/90—Stable introduction of foreign DNA into chromosome
- C12N15/902—Stable introduction of foreign DNA into chromosome using homologous recombination
- C12N15/907—Stable introduction of foreign DNA into chromosome using homologous recombination in mammalian cells
-
- A—HUMAN NECESSITIES
- A61—MEDICAL OR VETERINARY SCIENCE; HYGIENE
- A61K—PREPARATIONS FOR MEDICAL, DENTAL OR TOILETRY PURPOSES
- A61K38/00—Medicinal preparations containing peptides
- A61K38/16—Peptides having more than 20 amino acids; Gastrins; Somatostatins; Melanotropins; Derivatives thereof
-
- A—HUMAN NECESSITIES
- A61—MEDICAL OR VETERINARY SCIENCE; HYGIENE
- A61K—PREPARATIONS FOR MEDICAL, DENTAL OR TOILETRY PURPOSES
- A61K48/00—Medicinal preparations containing genetic material which is inserted into cells of the living body to treat genetic diseases; Gene therapy
-
- A—HUMAN NECESSITIES
- A61—MEDICAL OR VETERINARY SCIENCE; HYGIENE
- A61P—SPECIFIC THERAPEUTIC ACTIVITY OF CHEMICAL COMPOUNDS OR MEDICINAL PREPARATIONS
- A61P25/00—Drugs for disorders of the nervous system
-
- A—HUMAN NECESSITIES
- A61—MEDICAL OR VETERINARY SCIENCE; HYGIENE
- A61P—SPECIFIC THERAPEUTIC ACTIVITY OF CHEMICAL COMPOUNDS OR MEDICINAL PREPARATIONS
- A61P27/00—Drugs for disorders of the senses
- A61P27/02—Ophthalmic agents
-
- A—HUMAN NECESSITIES
- A61—MEDICAL OR VETERINARY SCIENCE; HYGIENE
- A61P—SPECIFIC THERAPEUTIC ACTIVITY OF CHEMICAL COMPOUNDS OR MEDICINAL PREPARATIONS
- A61P43/00—Drugs for specific purposes, not provided for in groups A61P1/00-A61P41/00
-
- C—CHEMISTRY; METALLURGY
- C07—ORGANIC CHEMISTRY
- C07K—PEPTIDES
- C07K14/00—Peptides having more than 20 amino acids; Gastrins; Somatostatins; Melanotropins; Derivatives thereof
- C07K14/195—Peptides having more than 20 amino acids; Gastrins; Somatostatins; Melanotropins; Derivatives thereof from bacteria
-
- C—CHEMISTRY; METALLURGY
- C12—BIOCHEMISTRY; BEER; SPIRITS; WINE; VINEGAR; MICROBIOLOGY; ENZYMOLOGY; MUTATION OR GENETIC ENGINEERING
- C12N—MICROORGANISMS OR ENZYMES; COMPOSITIONS THEREOF; PROPAGATING, PRESERVING, OR MAINTAINING MICROORGANISMS; MUTATION OR GENETIC ENGINEERING; CULTURE MEDIA
- C12N15/00—Mutation or genetic engineering; DNA or RNA concerning genetic engineering, vectors, e.g. plasmids, or their isolation, preparation or purification; Use of hosts therefor
-
- C—CHEMISTRY; METALLURGY
- C12—BIOCHEMISTRY; BEER; SPIRITS; WINE; VINEGAR; MICROBIOLOGY; ENZYMOLOGY; MUTATION OR GENETIC ENGINEERING
- C12N—MICROORGANISMS OR ENZYMES; COMPOSITIONS THEREOF; PROPAGATING, PRESERVING, OR MAINTAINING MICROORGANISMS; MUTATION OR GENETIC ENGINEERING; CULTURE MEDIA
- C12N15/00—Mutation or genetic engineering; DNA or RNA concerning genetic engineering, vectors, e.g. plasmids, or their isolation, preparation or purification; Use of hosts therefor
- C12N15/09—Recombinant DNA-technology
- C12N15/11—DNA or RNA fragments; Modified forms thereof; Non-coding nucleic acids having a biological activity
- C12N15/52—Genes encoding for enzymes or proenzymes
-
- C—CHEMISTRY; METALLURGY
- C12—BIOCHEMISTRY; BEER; SPIRITS; WINE; VINEGAR; MICROBIOLOGY; ENZYMOLOGY; MUTATION OR GENETIC ENGINEERING
- C12N—MICROORGANISMS OR ENZYMES; COMPOSITIONS THEREOF; PROPAGATING, PRESERVING, OR MAINTAINING MICROORGANISMS; MUTATION OR GENETIC ENGINEERING; CULTURE MEDIA
- C12N15/00—Mutation or genetic engineering; DNA or RNA concerning genetic engineering, vectors, e.g. plasmids, or their isolation, preparation or purification; Use of hosts therefor
- C12N15/09—Recombinant DNA-technology
- C12N15/11—DNA or RNA fragments; Modified forms thereof; Non-coding nucleic acids having a biological activity
- C12N15/62—DNA sequences coding for fusion proteins
-
- C—CHEMISTRY; METALLURGY
- C12—BIOCHEMISTRY; BEER; SPIRITS; WINE; VINEGAR; MICROBIOLOGY; ENZYMOLOGY; MUTATION OR GENETIC ENGINEERING
- C12N—MICROORGANISMS OR ENZYMES; COMPOSITIONS THEREOF; PROPAGATING, PRESERVING, OR MAINTAINING MICROORGANISMS; MUTATION OR GENETIC ENGINEERING; CULTURE MEDIA
- C12N15/00—Mutation or genetic engineering; DNA or RNA concerning genetic engineering, vectors, e.g. plasmids, or their isolation, preparation or purification; Use of hosts therefor
- C12N15/09—Recombinant DNA-technology
- C12N15/63—Introduction of foreign genetic material using vectors; Vectors; Use of hosts therefor; Regulation of expression
- C12N15/79—Vectors or expression systems specially adapted for eukaryotic hosts
- C12N15/82—Vectors or expression systems specially adapted for eukaryotic hosts for plant cells, e.g. plant artificial chromosomes (PACs)
- C12N15/8201—Methods for introducing genetic material into plant cells, e.g. DNA, RNA, stable or transient incorporation, tissue culture methods adapted for transformation
- C12N15/8213—Targeted insertion of genes into the plant genome by homologous recombination
-
- C—CHEMISTRY; METALLURGY
- C12—BIOCHEMISTRY; BEER; SPIRITS; WINE; VINEGAR; MICROBIOLOGY; ENZYMOLOGY; MUTATION OR GENETIC ENGINEERING
- C12N—MICROORGANISMS OR ENZYMES; COMPOSITIONS THEREOF; PROPAGATING, PRESERVING, OR MAINTAINING MICROORGANISMS; MUTATION OR GENETIC ENGINEERING; CULTURE MEDIA
- C12N15/00—Mutation or genetic engineering; DNA or RNA concerning genetic engineering, vectors, e.g. plasmids, or their isolation, preparation or purification; Use of hosts therefor
- C12N15/09—Recombinant DNA-technology
- C12N15/63—Introduction of foreign genetic material using vectors; Vectors; Use of hosts therefor; Regulation of expression
- C12N15/79—Vectors or expression systems specially adapted for eukaryotic hosts
- C12N15/85—Vectors or expression systems specially adapted for eukaryotic hosts for animal cells
-
- C—CHEMISTRY; METALLURGY
- C12—BIOCHEMISTRY; BEER; SPIRITS; WINE; VINEGAR; MICROBIOLOGY; ENZYMOLOGY; MUTATION OR GENETIC ENGINEERING
- C12N—MICROORGANISMS OR ENZYMES; COMPOSITIONS THEREOF; PROPAGATING, PRESERVING, OR MAINTAINING MICROORGANISMS; MUTATION OR GENETIC ENGINEERING; CULTURE MEDIA
- C12N9/00—Enzymes; Proenzymes; Compositions thereof; Processes for preparing, activating, inhibiting, separating or purifying enzymes
- C12N9/14—Hydrolases (3)
- C12N9/16—Hydrolases (3) acting on ester bonds (3.1)
- C12N9/22—Ribonucleases RNAses, DNAses
-
- C—CHEMISTRY; METALLURGY
- C12—BIOCHEMISTRY; BEER; SPIRITS; WINE; VINEGAR; MICROBIOLOGY; ENZYMOLOGY; MUTATION OR GENETIC ENGINEERING
- C12Y—ENZYMES
- C12Y301/00—Hydrolases acting on ester bonds (3.1)
- C12Y301/21—Endodeoxyribonucleases producing 5'-phosphomonoesters (3.1.21)
- C12Y301/21004—Type II site-specific deoxyribonuclease (3.1.21.4)
-
- A—HUMAN NECESSITIES
- A61—MEDICAL OR VETERINARY SCIENCE; HYGIENE
- A61K—PREPARATIONS FOR MEDICAL, DENTAL OR TOILETRY PURPOSES
- A61K38/00—Medicinal preparations containing peptides
-
- C—CHEMISTRY; METALLURGY
- C07—ORGANIC CHEMISTRY
- C07K—PEPTIDES
- C07K2319/00—Fusion polypeptide
- C07K2319/80—Fusion polypeptide containing a DNA binding domain, e.g. Lacl or Tet-repressor
Landscapes
- Health & Medical Sciences (AREA)
- Life Sciences & Earth Sciences (AREA)
- Genetics & Genomics (AREA)
- Engineering & Computer Science (AREA)
- Chemical & Material Sciences (AREA)
- Organic Chemistry (AREA)
- Bioinformatics & Cheminformatics (AREA)
- Biomedical Technology (AREA)
- Wood Science & Technology (AREA)
- Zoology (AREA)
- General Engineering & Computer Science (AREA)
- Biotechnology (AREA)
- Molecular Biology (AREA)
- General Health & Medical Sciences (AREA)
- Biochemistry (AREA)
- Microbiology (AREA)
- Biophysics (AREA)
- Plant Pathology (AREA)
- Physics & Mathematics (AREA)
- Medicinal Chemistry (AREA)
- Veterinary Medicine (AREA)
- Pharmacology & Pharmacy (AREA)
- Public Health (AREA)
- Animal Behavior & Ethology (AREA)
- Proteomics, Peptides & Aminoacids (AREA)
- General Chemical & Material Sciences (AREA)
- Nuclear Medicine, Radiotherapy & Molecular Imaging (AREA)
- Chemical Kinetics & Catalysis (AREA)
- Cell Biology (AREA)
- Gastroenterology & Hepatology (AREA)
- Epidemiology (AREA)
- Mycology (AREA)
- Neurosurgery (AREA)
- Neurology (AREA)
- Immunology (AREA)
- Ophthalmology & Optometry (AREA)
- Peptides Or Proteins (AREA)
- Micro-Organisms Or Cultivation Processes Thereof (AREA)
- Enzymes And Modification Thereof (AREA)
- Medicines That Contain Protein Lipid Enzymes And Other Medicines (AREA)
Abstract
Description
本出願は、2010年5月17日出願の米国仮出願第61/395,836号、2010年8月12日出願の同第61/401,429号、2010年10月13日出願の同第61/455,121号、2010年12月20日出願の同第61/459,891号、2011年2月2日出願の同第61/462,482号、2011年3月24日出願の同第61/465,869号の利益を主張し、それらの開示は、参照によりその全体が本明細書に組み込まれる。
該当なし
本発明は、遺伝子操作されたDNA結合タンパク質を使用した内在性遺伝子及び他のゲノム遺伝子座の発現状態の遺伝的修飾及び制御のための方法を提供する。
多くの、恐らくほとんどの生理学的及び病態生理学的プロセスを、遺伝子発現の選択的上方又は下方制御によって制御することができる。選択的制御によって制御することができる病理学の例として、数例を挙げると、リウマチ性関節炎における炎症性サイトカインの不適切な発現、高コレステロール血症における肝臓LDL受容体の過小発現、血管新生促進因子の過剰発現、及び固形腫瘍成長における抗血管新生因子の過小発現が挙げられる。加えて、ウイルス、細菌、真菌、及び原虫等の病原体を、それらの宿主細胞の遺伝子発現を変更することによって制御することができる。したがって、簡単に有益な遺伝子を上方制御し、かつ病原遺伝子を下方制御することができる治療的アプローチへの明らかに満たされていない必要性が存在する。
導入
本出願は、TALE反復ドメインを遺伝子操作して、所望の内因性DNA配列を認識することができること、及び機能ドメインのそのような遺伝子操作されたTALE反復ドメインへの融合を用いて、機能状態、又は遺伝子を含むその天然のクロマチン環境に存在する内因性細胞遺伝子座の実際のゲノムDNA配列を修飾することができることを実証する。したがって、本発明は、遺伝子を含む内因性細胞遺伝子座を高い有効性で特異的に認識するように遺伝子操作されたTALE融合DNA結合タンパク質を提供する。結果として、本発明のTALE融合物を、内在性遺伝子転写の活性化及び抑制の両方を介して、内在性遺伝子発現を制御するために使用することができる。TALE融合物を、他の制御又は機能ドメイン、例えば、ヌクレアーゼ、トランスポザーゼ、又はメチラーゼに結合して、内因性染色体の配列を修飾することもできる。
本方法の実践、並びに本明細書に開示の組成物の調製及び使用は、別途示されない限り、当技術分野の技術の範囲内の分子生物学、生化学、クロマチン構造及び分析、計算化学、細胞培養、組換えDNA、並びに関連分野における従来の技法を採用する。これらの技法は、文献において十分に説明されている。例えば、Sambrook et al.MOLECULAR CLONING:A LABORATORY MANUAL,Second edition,Cold Spring Harbor Laboratory Press,1989 and Third edition,2001、Ausubel et al.,CURRENT PROTOCOLS IN MOLECULAR BIOLOGY,John Wiley & Sons,New York,1987及び定期的に更新されたもの、METHODS IN ENZYMOLOGYシリーズ,Academic Press,San Diego、Wolffe,CHROMATIN STRUCTURE AND FUNCTION,Third edition,Academic Press,San Diego,1998、METHODS IN ENZYMOLOGY,Vol.304,“Chromatin”(P.M.Wassarman and A.P.Wolffe,eds.),Academic Press,San Diego,1999、並びにMETHODS IN MOLECULAR BIOLOGY,Vol.119,“Chromatin Protocols”(P.B.Becker,ed.)Humana Press,Totowa,1999を参照されたい。
「核酸」、「ポリヌクレオチド」、及び「オリゴヌクレオチド」という用語は、同義に使用され、線状若しくは環状立体配座、かつ一本鎖形態若しくは二本鎖形態のいずれかのデオキシリボヌクレオチド又はリボヌクレオチドポリマーを指す。本開示の目的において、これらの用語は、ポリマーの長さを限定すると解釈されるべきではない。本用語は、天然ヌクレオチドの既知の類似体、並びに塩基、糖、及び/又はリン酸塩部分(例えば、ホスホロチオエートバックボーン)で修飾されるヌクレオチドを網羅し得る。概して、特定のヌクレオチドの類似体は、同一の塩基対形成特異性を有し、すなわち、Aの類似体は、Tと塩基対を形成する。
エピソームの例には、プラスミド及びある特定のウイルスゲノムが挙げられる。
本明細書に記載のポリペプチドは、1つ以上(例えば、1、2、3、4、5、6、7、8、9、10、11、12、13、14、15、16、17、18、19、20個以上)のTALE反復単位を含む。複数のTALE反復単位を含むTALE DNA結合ドメインが、特異性に関与する配列を決定するために研究されている。1つの生物内で、TALE反復は、典型的には、高度に保存される(RVDを除く)が、異なる種にわたってはよく保存されない場合もある。
本明細書に記載のDNA結合タンパク質を含む融合タンパク質(例えば、TALE融合タンパク質)及び異種制御若しくは機能ドメイン(又はその機能フラグメント)も提供される。一般的なドメインには、例えば、転写因子ドメイン(活性化因子、抑制因子、共活性化因子、共抑制因子)、ヌクレアーゼドメイン、サイレンサードメイン、癌遺伝子ドメイン(例えば、myc、jun、fos、myb、max、mad、rel、ets、bcl、myb、mosファミリーメンバー等);DNA修復酵素並びにそれらの関連因子及び修飾因子;DNA転位酵素並びにそれらの関連因子及び修飾因子;クロマチン関連タンパク質並びにそれらの修飾因子(例えば、キナーゼ、アセチラーゼ、及びデアセチラーゼ);並びにDNA修飾酵素(例えば、メチルトランスフェラーゼ、トポイソメラーゼ、ヘリカーゼ、リガーゼ、キナーゼ、ホスファターゼ、ポリメラーゼ、エンドヌクレアーゼ)、DNA標的酵素、例えば、トランスポゾン、インテグラーゼ、リコンビナーゼ、及びリゾルバーゼ、並びにそれらの関連因子及び修飾因子、核ホルモン受容体、ヌクレアーゼ(切断ドメイン又は半ドメイン)、及びリガンド結合ドメインが含まれる。他の融合タンパク質は、レポーター又は選択マーカーを含み得る。レポータードメインの例には、GFP、GUS等が挙げられる。植物細胞において特定の有用性を有するレポーターは、GUSを含む。
任意の所望の遺伝子(複数を含む)中に標的部位を含む任意のヌクレアーゼを、本明細書に開示の方法において使用することができる。例えば、ホーミングエンドヌクレアーゼ及びメガヌクレアーゼは、非常に長い認識配列を有し、それらのうちのいくつかは、統計的基礎に従って、ヒト大のゲノムに一度存在する可能性が高い。所望の遺伝子中に標的部位を有する任意のそのようなヌクレアーゼを、標的化切断のために、例えば、亜鉛フィンガーヌクレアーゼ及び/若しくはメガヌクレアーゼを含むTALE反復ドメインヌクレアーゼ融合物の代わりに、又はそれに加えて使用することができる。
TALE融合タンパク質、TALE融合タンパク質をコードするポリヌクレオチド、並びに本明細書に記載のタンパク質及び/又はポリヌクレオチドを含む組成物を、例えば、TAL融合タンパク質をコードするmRNAの注入を含む、任意の好適な手段によって、標的細胞に送達することができる。Hammerschmidt et al.(1999)Methods Cell Biol.59:87−115を参照されたい。
本明細書に記載の方法及び組成物は、植物、動物(例えば、マウス、ラット、霊長類、家畜、ウサギ等の哺乳動物)、魚等の真核生物を含むが、それらに限定されない、遺伝子発現を制御し、かつ/又はゲノム修飾を介して生物を変更することが所望される任意の生物に適用できる。真核(例えば、酵母、植物、真菌、魚、並びに猫、犬、マウス、ウシ、羊、及びブタ等の哺乳類細胞)細胞を使用することができる。本明細書に記載の1つ以上のホモ接合KO遺伝子座又は他の遺伝的修飾を含有する生物由来の細胞も使用することができる。
様々なアッセイを使用して、TALE融合タンパク質による遺伝子発現制御のレベルを決定することができる。特定のTALE融合タンパク質の活性を、様々なインビトロ及びインビボアッセイを使用して、例えば、タンパク質又はmRNAレベル、産物レベル、酵素活性、腫瘍成長の測定によって、レポーター遺伝子の転写活性化又は抑制によって、第2のメッセンジャーレベル(例えば、cGMP、cAMP、IP3、DAG、Ca.sup.2+)によって、サイトカイン及びホルモン産生レベルによって、並びに例えば、免疫アッセイ(例えば、抗体を用いたELISA及び免疫組織化学的アッセイ)、ハイブリダイゼーションアッセイ(例えば、RNase保護、ノーザンブロット、in situハイブリダイゼーション、オリゴヌクレオチド配列研究)、比色アッセイ、増幅アッセイ、酵素活性アッセイ、腫瘍成長アッセイ、表現型アッセイ等を用いた新血管形成によって評価することができる。
従来のウイルス及び非ウイルスに基づく遺伝子導入方法を使用して、全生物又は標的組織中の哺乳類細胞において遺伝子操作されたTALEドメイン融合物をコードする核酸を導入することができる。そのような方法を使用して、インビトロで、TALEドメイン融合物をコードする核酸を細胞に投与することができる。好ましくは、TALEドメイン融合物をコードする核酸は、インビボ又はエクスビボ使用のために投与される。非ウイルスベクター送達系は、DNAプラスミド、ネイキッド核酸、及びリポソーム等の送達ビヒクルと錯体形成された核酸を含む。ウイルスベクター送達系は、細胞への送達後にエピソームゲノム又は組込みゲノムのいずれかを有するDNA及びRNAウイルスを含む。遺伝子治療手順の総説については、Anderson,Science 256:808−813(1992)、Nabel & Felgner,TIBTECH 11:211−217(1993)、Mitani & Caskey,TIBTECH 11:162−166(1993)、Dillon,TIBTECH 11:167−175(1993)、Miller,Nature 357:455−460(1992)、Van Brunt,Biothechnology 6(10):1149−1154(1988);Vigne,Restorative Neurology and Neuroscience 8:35−36(1995)、Kremer & Perricaudet,British Medical Bulletin 51(1):31−44(1995)、Haddada et al.,in Current Topics in Microbiology and Immunology Doerfler and Bohm(eds)(1995)、及びYu et al.,Gene Therapy 1:13−26(1994)を参照されたい。
TALE融合物及びTALE融合物をコードする発現ベクターを、遺伝子発現の調節のため、かつ治療的又は予防的用途、例えば、癌、虚血、糖尿病性網膜症、黄斑変性、リウマチ性関節炎、乾癬、HIV感染、鎌状赤血球貧血、アルツハイマー病、筋ジストロフィー、神経変性疾患、血管疾患、嚢胞性線維症、脳卒中等のために、患者に直接投与することができる。TALE融合タンパク質遺伝子治療によって阻害することができる微生物の例には、病原菌、例えば、クラミジア、リケッチア細菌、マイコバクテリア、ブドウ球菌、連鎖球菌、肺炎球菌、髄膜炎菌及び淋菌、クレブシエラ菌、プロテウス菌、セラチア菌、シュードモナス菌、レジオネラ菌、ジフテリア、サルモネラ菌、桿菌、コレラ菌、破傷風、ボツリヌス中毒症、炭疽菌、ペスト、レプトスピラ症、及びライム病原因菌;感染性真菌、例えば、アスペルギルス属、カンジダ種;胞子虫類(例えば、マラリア原虫)、根足虫(例えば、エントアメーバ属)、及び鞭毛虫(トリパノソーマ属、リーシュマニア属、トリコモナス属、ジアルジア属等)等の原虫;ウイルス疾患、例えば、肝炎(A型、B型、又はC型)、ヘルペスウイルス(例えば、VZV、HSV−1、HSV−6、HSV−II、CMV、及びEBV)、HIV、エボラ、アデノウイルス、インフルエンザウイルス、フラビウイルス、エコーウイルス、ライノウイルス、コクサッキーウイルス、コモウイルス、呼吸器合胞体ウイルス、ムンプスウイルス、ロタウイルス、麻疹ウイルス、風疹ウイルス、パルボウイルス、ワクチニアウイルス、HTLVウイルス、デング熱ウイルス、乳頭腫ウイルス、ポリオウイルス、狂犬病ウイルス、及びアルボウイルス、脳炎ウイルス等が挙げられる。
TALE融合物を使用して、病害抵抗性の増加、構造及び貯蔵多糖類、風味、タンパク質、及び脂肪酸の修飾、果実の熟成、収率、色、栄養学的特徴、貯蔵能力の向上、渇水又は冠水/浸水耐性等の形質のために、植物を遺伝子操作することができる。具体的には、油産生の強化のための作物種の遺伝子操作、例えば、油料種子において産生される脂肪酸の修飾を目的とする。例えば、米国特許第7,262,054号、並びに米国特許公開第2008/0182332号及び同第20090205083号を参照されたい。
TALE融合物は、遺伝子発現の表現型結果及び機能を決定するアッセイのための用途を有する。分析技法における近年の進歩は、集中的な大量の配列決定の試みと相まって、以前に利用可能であった分子標的よりもはるかに多くの分子標的を同定及び特性化する機会をもたらしている。遺伝子及びそれらの機能についてのこの新しい情報は、基本的な生物学的理解を加速し、治療的介入のために多くの新しい標的を提示する。いくつかの場合において、分析ツールは、新しいデータの生成に追いついていない。全体的な差次的遺伝子発現の測定における近年の進歩による例が提供される。遺伝子発現マイクロアレイ、差次的cDNAクローニング頻度、減算的ハイブリダイゼーション、及びディファレンシャルディスプレイ法に代表されるこれらの方法は、異なる組織内で、又は特定の刺激物に応答して、上方若しくは下方制御される遺伝子を非常に迅速に同定することができる。そのような方法は、形質転換、腫瘍進行、炎症応答、神経障害等の生物学的プロセスを調査するためにますます使用されている。当業者は、現在、所与の生理学的現象と相関する差次的発現遺伝子の長いリストを非常に容易に生成することができるが、個々の差次的発現遺伝子と現象との間の因果関係を実証するのは困難である。現在のところ、機能を差次的発現遺伝子に割り当てるための簡単な方法は、差次的遺伝子発現を監視する能力に追いついていない。
TALE融合技術のさらなる適用は、遺伝子発現を操作し、かつ/又はゲノムを変更して、トランスジェニック動物又は植物を産生する。細胞株と同様に、内在性遺伝子の過剰発現又はトランスジェニックマウス等のトランスジェニック動物への異種遺伝子の導入は、極めて容易なプロセスである。同様に、トランスジェニック植物の産生は周知である。本明細書に記載のTALEドメイン融合技術を使用して、トランスジェニック動物及び植物を容易に生成することができる。
開示の方法及び組成物を使用して、所望の遺伝子座で遺伝子制御を制御することができる。最適な遺伝子を、TALE反復ドメインに融合する転写制御ドメインに応じて、活性化するか、又は抑制することができる。TALE活性化因子を、分化した細胞からiPSCを産生するという目標のために、多能性誘導遺伝子に標的化することができる。これは、特定の病状のためのインビトロ及びインビボモデル開発に、かつiPSCに由来する細胞治療薬の開発に役立ち得る。
初期の設計フレームワークとしての機能を果たし得る天然TALEタンパク質を同定するために、高度の特異性、並びに哺乳類細胞における標的配列結合の証拠の両方を呈する基準の天然TALEを同定した。具体的には、12.5個のTALE反復(TALE13と称される、12個の全反復及び1個の半反復)を含有するTALEタンパク質を、以下のプライマー対:pthA d152N EcoR、ACGTGGATTCATGGTGGATCTACGCACGCTC(配列番号52)及びpthA Sac2 Rev、TACGTCCGCGGTCCTGAGGCAATAGCTCCATCA(配列番号53)を使用して、PCR増幅によって、キサントモナス・アクソノポディスからクローニングした。プライマー対を、最初は、N末端の152個のアミノ酸を切断したAvrBs3遺伝子を増幅するように設計した。これらの配列は、植物細胞への輸送に必要であるが、さもなければ機能にとっては不必要であることが以前に示されている(Szurek et al(2002)Mol.Micro 46(1)p.13−23を参照のこと)。中心タンデム反復の数の変化を有する高度に保存された配列を特徴とするいくつかのTALEタンパク質を、これらのプライマー対を用いてPCRによって単離した。hssB3.0として報告されたTALE15を除いて(Shiotani et al(2007)J.Bacteriol 189(8):3271−9)、単離した他のTALEタンパク質は、公開文献において報告されていないため、新規のタンパク質のようである。これらには、それぞれ、13、9、及び16個のTALE反復を有する、TALE13、TALE9、及びTALE16が含まれる。
最大活性を提供するキャッピング配列の範囲の初期の調査として、いくつかのTALE切断を行った。これらの切断が、以下の表4に示される。
N18TA:5’CAGGGATCCATGCACTGTACGTTTNNNNNNNNNNNNNNNNNNAAACCACTTGACTGCGGATCCTGG3’(配列番号55)を有するDNA二本鎖を含み、Nは、4つ全ての塩基の混合物を示す。さらなるライブラリ(示されるような)は、以下の配列を含む:
N22AT:5’CAGGGATCCATGCACTGTACGAAANNNNNNNNNNNNNNNNNNNNNNTTTCCACTTGACTGCGGATCCTGG3’(配列番号59)
N21TA:5’CAGGGATCCATGCACTGTACGTTTNNNNNNNNNNNNNNNNNNNNNAAACCACTTGACTGCGGATCCTGG3’(配列番号60)
N23TA:5’CAGGGATCCATGCACTGTACGTTTNNNNNNNNNNNNNNNNNNNNNNNAAACCACTTGACTGCGGATCCTGG3’(配列番号61)
N26:5’CAGGGATCCATGCACTGTACGTTNNNNNNNNNNNNNNNNNNNNNNNNNNAACCACTTGACTGCGGATCCTGG3’
N30CG:
5’CAGGGATCCATGCACTGTACGCCCNNNNNNNNNNNNNNNNNNNNNNNNNNNNNNGGGCCACTTGACTGCGGATCCTGG3’(配列番号62)
2つのさらなる天然TALEタンパク質を、SELEX手順に供して、これらのタンパク質が結合する標的DNA配列を同定した。TALE9が、以下のDNA標的:TANAAACCTT(配列番号56)を特定する8.5個のTALE反復を有する一方で、TALE16は、以下の標的:TACACATCTTTAACACT(配列番号57)を予測する15.5個のTALE反復を有する。データは、表9及び10に示される。表9において、クローン番号2の構造のTALE9タンパク質が使用されており、結果が示されている。TALE13、クローン番号2と同様に、この実験は、第2の部分的にランダム化されたDNAライブラリで繰り返され、第1のライブラリと類似のデータをもたらした。TALE13について上で記載したように、TALE9は、その標的配列に高度に特異的である。
哺乳類細胞におけるTALEドメイン融合物の機能活性を調査するために、遺伝子操作されたレポーター構築物を以下のように作製した。クローニングされたTALE13又はTALE15の標的配列の1つ以上のコピーを、NheI部位とBglII部位との間のレポーター構築物に挿入し、それによって、標的を、pGL3プラスミド(Promega)中の最小のSV40プロモーターによって駆動されるホタルルシフェラーゼ発現単位から上流に配置する(図2を参照のこと)。pGL3プラスミドのプロモーター領域が図2Aに示され、TALE13の2つの予測標的部位を含有する配列が、図2Bに示される。図3に示される実験において、TALEタンパク質構築物(2つの標的を含有するレポータープラスミドとともに)(図3A)、及び内部対照としてレニラルシフェラーゼ(Promega)を含有する発現構築物を、ヒト293細胞に共トランスフェクトした。その後、それぞれのTALEタンパク質によって誘導されたホタルルシフェラーゼ活性を、トランスフェクションの2日後に分析した。複数の標的に応じて、TALE VP16融合物は、哺乳類細胞におけるレポーター遺伝子発現を相乗的に活性化することができる(図3)。さらに、図4Bに示されるように、VP16活性化ドメイン(TR13−VP16及びTR15−VP16)を付加したTALEタンパク質は、ルシフェラーゼレポーター遺伝子を活性化する。VP16ドメインを有しない天然TALEタンパク質の発現は、ルシフェラーゼを活性化しない(TR13及びTR15)。したがって、レポーター遺伝子活性化は、正しい標的がそれらの対応するTALE融合物と一致する場合にのみ観察され、転写活性化が標的化DNA結合に起因することを示唆する。
TALEタンパク質を転写制御ドメインに結合して、哺乳類細胞におけるレポーター遺伝子発現を調節することができることを実証した後、所望の標的特異性を有するTALE転写因子を遺伝子操作するための実験を行った。TR13 VP16のサイレント変異(すなわち、アミノ酸配列の変更を伴わないヌクレオチド配列の変更)を導入して、それぞれ、最初のタンデム反復の開始部及び最後のタンデム反復の最後部に、2つの特有の制限部位、ApaI及びHpaIを作成した。その後、これらのApaI及びHpaI部位を、合成タンデム反復をTR13 VP16骨格にクローニングするために使用し、タンデム反復に隣接する完全なN末端及びC末端配列、並びにVP16活性化ドメインを有する遺伝子操作されたTALEを生成した。
この実施例では、以下の表15に示されるように、TALE13タンパク質N末端又はC末端での種々の欠失を生成した。
次に、人工TALEヌクレアーゼ(TALEN)に関連したTALEのDNA標的化能力を評価した。FokIドメイン間でGGGS配列の12個のコピーによって結合されるFokIヌクレアーゼドメインの2つのコピーが、VP16活性化ドメインを置換するために使用されたことを除いて、上述のR13V−d182Cと同一であるR13d182C−scFokIと称される構築物を生成するために、実施例6で定義されるTALE13のDNA標的ドメインを、ヌクレアーゼドメインに結合させた。次に、TALEN構築物を、一本鎖アニーリング(SSA)に基づくレポーターアッセイにおけるヌクレアーゼ活性について試験した(共同所有の米国特許公開第20110014616号を参照のこと)。
上述の二量体対を、NTF3遺伝子座に標的化し(表18を参照のこと)、その後、哺乳類細胞中の内因性遺伝子座で試験した。示される二量体対を、製造業者によって提供される標準の方法を用いて、Amaxa Biosystemsデバイス(Cologne,Germany)を使用してK562細胞にヌクレオフェクトし、トランスフェクション後に一過性の低温ショック成長条件に供した(米国出願第12/800,599号を参照のこと)。
NTF3でのTALE媒介標的組込みは、HDR DNA修復経路を介して、又はNHEJ経路を介して起こり得る。我々は、NHEJによる小さい二本鎖オリゴヌクレオチドの捕捉に基づいてNTF3でのTALE媒介標的組込みをアッセイする実験を設計した。我々は、先に、ZFN誘導DNA二本鎖切断(DSB)の部位でのオリゴヌクレオチドの捕捉を示した。この種の標的組込みは、ZFN対のFokI部分によって作成されるものに相補的な5’オーバーハングの存在によって強化された(が、絶対に必要ではなかった)。FokIは、4bpの5’オーバーハングを自然に作成し、ZFNとの関連で、FokIヌクレアーゼドメインは、4bp又は5bpのいずれかの5’オーバーハングを作成する。NTF3 TALENによって残されるオーバーハングの位置及び組成が未知であるため、我々は、NTF3 TALEN結合部位(NT3−1F〜NT3−9R)の間の12bpのスペーサー領域における全ての可能性のある4bpの5’オーバーハングで9個の二本鎖オリゴヌクレオチドドナーを設計した(表20を参照のこと)。
天然タンパク質において見出されるTALE反復をコードするDNA配列は、それらの対応するアミノ酸配列と同程度に繰り返される。天然TALEは、典型的には、それぞれの反復の配列間に数個のみの塩基対分の相違を有する。反復DNA配列は、所望の全長DNA増幅産物を効率的に増幅することを困難にし得る。これは、天然TALEを含有するタンパク質のためにDNAを増幅することを試みるときに示されている。Mfold(M.Zuker,Nucleic Acids Res.31(13):3406−15,(2003))を使用した上のTALE反復タンパク質のDNA配列のさらなる分析は、それらが、効率的な増幅を破壊する反復配列を有するのみならず、非常に安定した二次構造も含有することを明らかにした。この分析において、第1の全反復配列をコードする核酸の5’末端で開始する配列の800個の塩基対を分析した。したがって、分析した核酸配列は、約7.5反復の配列を含有した。これらの二次構造のうちのいくつかが、図21に示される。
様々なTALE融合タンパク質の迅速な組み立てを可能にするために、ともに結合することができる反復モジュールのアーカイブを作成して、ほぼ任意の選択される標的DNA配列に特異的なTALE DNA結合ドメインを作成する方法を開発した。所望の標的DNA配列に基づいて、1つ以上のモジュールを選択し、PCRに基づくアプローチを介して取り出す。モジュールを、タンデムに結合し、最適な融合パートナードメインを含有するベクター骨格の中に連結する。
プライマー:
T1F−Bsa GGATCCGGATGGTCTCAACCTGACCCCAGACCAG(配列番号106)
T1R−Bsa GAGGGATGCGGGTCTCTGAGTCCATGATCCTGGCACAGT(配列番号107)
T2F−Bsa GGATCCGGATGGGTCTCAACTCACCCCAGACCAGGTA(配列番号108)
T2R−Bsa GAGGGATGCGGGTCTCTCAGCCCATGATCCTGGCACAGT(配列番号109)
T3F−Bsa GGATCCGGATGGGTCTCAGCTGACCCCAGACCAG(配列番号110)
T3R−Bsa GAGGGATGCGGGTCTCTCAAACCATGATCCTGGCACAGT(配列番号111)
T4F−Bsa GGATCCGGATGGGTCTCATTTGACCCCAGACCAGGTA(配列番号112)
T4R−Bsa CTCGAGGGATGGTCTCCTGTCAGGCCATGATCC(配列番号113)
TALEN設計方法を評価するために、我々は、ヒトCCR5遺伝子内のデルタ32変異(以下に太字下線で示される)の位置付近でのTALEN媒介遺伝子修飾の実証に努めた(Stephens JC et al,(1998)Am J Hum Gen 62(6):1507−15を参照のこと)。この研究のために、我々は、16個の二量体標的(配列番号114〜122)のパネルを定義した、デルタ32の位置で4つの「左」及び4つの「右」結合部位(以下を参照のこと)のクラスターを指定した。
2つの好ましいTALEN構造(C+28Cキャップ又はC+63Cキャップ対)のギャップ間隔選好を試験するために、ギャップ間隔に従って、C+28/C+28又はC+63/C+63の対形成を含有する全てのTALEN対を活性について分類した。結果が図25に示され、小さい方のTALENタンパク質、C+28/C+28対が、より制約されたギャップ間隔選好を有し、標的配列が12個又は13個の塩基対のギャップで分離される標的において最も高い活性を示すことを実証する。逆に、図25Bに示される大きい方のTALENタンパク質、C+63/C+63対は、12〜23個の塩基対のギャップ間隔を含有する標的において活性を示す。
ヌクレアーゼドメインを高度に活性なヌクレアーゼ機能を提供するTALE反復配列に結合させるために使用することができる組成物の体系的マッピング。最初に、1つのTALEN対を、2つの結合ドメイン間の定義されたギャップ間隔を有する単一の標的に対して選択した。選択したTALEN対は、CCR5遺伝子に特異的であり、かつ18個の塩基対ギャップ間隔を有する、L538/R557対として実施例12に記載されるものであった。一連の切断がC−2〜C+278のCキャップをもたらすように、欠失を上述のように行った。
代替の(非定型)RVDを調査して、DNA結合特異性を決定する位置の他のアミノ酸を変更することができるかを決定した。その結合活性は、SELEX及びELISAによって中央位置でのミスマッチに敏感であると示されたTALE結合ドメインを構築した。このタンパク質は、配列5’−TTGACAATCCT−3’(配列番号178)に結合し、配列5’−TTGACCATCCT−3’(配列番号179)、5’−TTGACGATCCT−3’(配列番号180)、又は5’−TTGACTATCCT−3’(配列番号181)に対してわずかな結合活性示した(図27に示されるELISAデータ)。これらの標的は、中間の三重核酸を意味するCXA標的と称され、Xは、A、C、T、又はGのいずれかである。
A:RI、KI、HI
C:ND、KD、AD
G:DH、SN、AK、AN、DK、HN
T:VG、IA、IP、TP、QA、YG、LA、SG、HA、NA、GG、KG、QG
N:KS、AT、KT、RA
天然TALEの大部分は、Tヌクレオチド塩基との相互作用を特定するために、C末端半反復中のNG RVDを使用する。したがって、新規のC末端半反復の生成を調査して、TALE標的の拡大を可能にした。Pou5F1及びPITX3遺伝子を標的とするTALENを、バックボーンとして使用し、C末端半反復内のRVD(Cキャップアミノ酸C−9及びC−8)を変更して、代替の核酸を特定した。これらの変異体中に、Aを認識するためにNI RVDを、Cを認識するためにHDを、Gを認識するためにNKを挿入し、対照は、Tを認識するためのNGであった。使用したTALENは、15〜18個のRVDを含有し、これらの2つの遺伝子中の様々な標的配列を標的とした。
最適な標的配列、ひいては最適なTALENタンパク質設計を決定するために、複数のSELEXアッセイからの結果を使用してインシリコ分析を行い、i)R1反復(N末端反復)単位の最良の標的、かつii)二量体及び三量体環境において、それらの隣接する反復単位と関連して、特定のRVD反復がどのように挙動するかを決定した。これらの研究において、Aを認識するためにNI RVDを、Cを認識するためにHDを、Gを認識するためにNNを、Tを認識するためにNGを使用した。
TALEN系の多用途性を実証するために、TALENを用いて、ヒト胚幹細胞(ESC)及び誘導性多能性幹細胞(iPSC)における標的組込みを駆動した。ヒトESC及びiPSCを、ピューロマイシンマーカーの発現がAAVS1プロモーターによって駆動されるAAVS1遺伝子座への制限部位をさらに含むピューロマイシンドナー核酸の標的組込みために使用した。ドナー及び従った方法は、共同所有の国際公開第WO2010117464号において以前に説明されているドナー及び方法であった(Hockemeyer et al(2009)Nat Biotechnol 27(9):851−857も参照されたく、その中で、我々は、AAVS1遺伝子座へのそのような構築物の標的組込みの自発的頻度が、我々のアッセイの検出の限界より低いことを実証した)。使用したヌクレアーゼは、実施例11に記載のAAVS1遺伝子座に特異的なTALENであり、標的結合部位は、以下に示される:
101125:GACCCTGCCTGCTCCT(配列番号329)
101225:CACCTGCAGCTGCCCAG(配列番号330)
TALENは、+63Cキャップを利用し、典型的なRVD(それぞれ、A、C、G、及びTを標的とする、NI、HD、NN、及びNG)を使用した。101125は、15.5個のTALE反復を含み、101225は、16.5個のTALE反復を含んだ。101225は、その標的部位における3’Gを認識するために、NN RVDを有する半反復を利用した。
101148:GGCCCTTGCAGCCGT(配列番号331)
101146:CAGACGCTGGCACT(配列番号332)
カエノラブディティス・エレガンスにおけるTALENゲノム編集。インビボ遺伝子編集のためにTALENを動物において使用することができることを実証するために、以下の実験を行った。カエノラブディティス・エレガンスben−1変異に特異的なTALEN対を、RNAとして送達し、Driscoll et al((1989)J.Cell.Biol.109:2993−3003)に記載されるベノミル耐性についてスクリーニングした。ben−1変異体表現型は優性であり、通常の解剖顕微鏡下で子孫の100%において可視的である。簡潔に、野生型カエノラブディティス・エレガンス雌雄同体を、ben−1を標的とするTALENをコードするmRNAの注入前に、通常のNGM寒天プレート上で栽培した。
GJC 153Fプライマー:5’ggaggcaagaagatggattc(配列番号453)
GJC 154Rプライマー:5’gaatcggcacatgcagatct(配列番号454)
TALE反復単位における変更を調査するために、キサントモナス及びラルストニアの両方からの配列を比較した。ラルストニア由来の52個の特有の反復単位を試験して、それぞれの位置における残基頻度を観察し、その後、これらの値を編集した。データが以下の表38に示され、アミノ酸は、左から右に1文字のコードで示され、反復単位の位置は、上から下に示され、RVD位置は、太字で示される。
亜鉛フィンガーをTALE DNA結合ドメインに融合させて、ハイブリッドDNA結合ドメインを作成し、その後、ヌクレアーゼに結合させた。標的DNA配列が以下に示され、CCR5遺伝子内に遺伝子座を取り囲む領域を含む。結合部位の上及び下に示されるのは、TALE DNA結合ドメインの標的結合部位であり、亜鉛フィンガー結合部位は、標的配列上に太字下線で示される。太字/下線の「TAG」配列が、CCR5特異的ZFN SBS番号8267の第4のフィンガーの結合部位である一方で、太字/下線の「AAACTG」配列は、CCR5特異的ZFN SBS番号8196の第3及び第4のフィンガーの結合部位である(米国特許出願第11/805,707号を参照のこと)。以下の配列は、亜鉛フィンガーDNA標的が、DNA鎖上のTALE DNA結合ドメイン標的と隣接せず、「内部ギャップ」を作成することを示す。したがって、この種の融合は、実践者が、所望の場合、内部ギャップ領域内でDNA領域を飛ばすことを可能にする。
ある特定のホットスポットへの選好が存在するが、レトロウイルスのライフサイクルの間、ウイルスゲノムRNAを逆転写し、かつ多くの異なる部位で宿主ゲノムに組み込む。レトロウイルスベクターを利用する適用、特に遺伝子治療において、癌遺伝子座付近での遺伝子操作されたウイルスゲノムのランダム組込みによるレトロウイルスベクターの考えられる発癌性は、潜在的な危険因子を示す。そのような潜在的な問題を打開するために、特定のTALE DNA結合ドメインを利用することによって、ウイルスインテグラーゼの特異性を既定の部位に再指向する。融合物を、全体若しくは切断されたインテグラーゼ、及び全体若しくは切断されたインテグラーゼ結合タンパク質で作製する(例えば、HIVインテグラーゼのLEDGF)。さらに、対の1つのメンバーが1つのタンパク質(例えば、タンパク質1)に融合する組込み体であり、第2の対がTALE DNA結合ドメインの別のタンパク質(例えば、タンパク質2)との融合物である融合対を作製し、タンパク質1及びタンパク質2は相互に結合する。対が目的とする細胞中で発現されるように、融合対を発現ベクターにクローニングする。哺乳類ゲノム標的について、融合対は、発現哺乳類発現ベクターを使用して発現される。TALEN誘導性DNA融合後にドナーが切断部位に組み込まれるように、TALEN融合物の発現の間、ドナーDNAを供給する。
DNA及びタンパク質配列
GACTCTTCGCGATGTACGGGCCAGATATACGCGTTGACATTGATTATTGACTAGTTATTAATAGTAATCAATTACGGGGTCATTAGTTCATAGCCCATATATGGAGTTCCGCGTTACATAACTTACGGTAAATGGCCCGCCTGGCTGACCGCCCAACGACCCCCGCCCATTGACGTCAATAATGACGTATGTTCCCATAGTAACGCCAATAGGGACTTTCCATTGACGTCAATGGGTGGAGTATTTACGGTAAACTGCCCACTTGGCAGTACATCAAGTGTATCATATGCCAAGTACGCCCCCTATTGACGTCAATGACGGTAAATGGCCCGCCTGGCATTATGCCCAGTACATGACCTTATGGGACTTTCCTACTTGGCAGTACATCTACGTATTAGTCATCGCTATTACCATGGTGATGCGGTTTTGGCAGTACATCAATGGGCGTGGATAGCGGTTTGACTCACGGGGATTTCCAAGTCTCCACCCCATTGACGTCAATGGGAGTTTGTTTTGGCACCAAAATCAACGGGACTTTCCAAAATGTCGTAACAACTCCGCCCCATTGACGCAAATGGGCGGTAGGCGTGTACGGTGGGAGGTCTATATAAGCAGAGCTCTCTGGCTAACTAGAGAACCCACTGCTTACTGGCTTATCGAAATTAATACGACTCACTATAGGGAGACCCAAGCTGGCTAGCGTTTAAACTTAAGCTGATCCACTAGTCCAGTGTGGTGGAATTCGCCATGGACTACAAAGACCATGACGGTGATTATAAAGATCATGACATCGATTACAAGGATGACGATGACAAGATGGCCCCCAAGAAGAAGAGGAAGGTGGGCATTCACGGGGTACCCGCCGCTGTGGATCTACGCACGCTCGGCTACAGCCAGCAGCAACAGGAGAAGATCAAACCGAAGGTTCGTTCGACAGTGGCGCAGCACCACGAGGCACTGGTCGGCCATGGGTTTACACACGCGCACATCGTTGCGCTCAGCCAACACCCGGCAGCGTTAGGGACCGTCGCTGTCAAGTATCAGGACATGATCGCAGCGTTGCCAGAGGCGACACACGAAGCGATCGTTGGCGTCGGCAAACAGTGGTCCGGCGCACGCGCCCTGGAGGCCTTGCTCACGGTGGCGGGAGAGTTGAGAGGTCCACCGTTACAGTTGGACACAGGCCAACTTCTCAAGATTGCAAAACGTGGCGGCGTGACCGCAGTGGAGGCAGTGCATGCATGGCGCAATGCACTGACGGGGGCCCCCCTGAACCTTACTCCCGAACAAGTAGTAGCGATAGCCAGTAATAACGGAGGTAAACAAGCCTTGGAGACGGTCCAAAGGTTGCTCCCGGTCTTGTGTCAGGCACATGGGCTGACGCCTCAACAGGTCGTCGCGATAGCGTCTAATAATGGAGGAAAGCAAGCTCTGGAAACCGTCCAGCGACTCCTTCCGGTTCTGTGCCAGGCTCATGGTCTGACTCCGCAGCAAGTCGTTGCTATAGCGTCCAACATCGGAGGCAAACAGGCCCTGGAGACCGTGCAGCGGTTGTTGCCTGTGCTTTGCCAAGCCCACGGGCTTACGCCTGAGCAAGTGGTGGCGATTGCCAGTAACAACGGCGGCAAACAAGCCCTTGAGACTGTGCAGAGGCTCTTGCCGGTACTCTGCCAAGCACACGGCTTGACCCCCGAGCAGGTTGTAGCCATAGCTAGTCACGACGGGGGTAAGCAAGCGTTGGAAACGGTGCAAGCACTTCTCCCCGTTCTCTGTCAAGCGCATGGACTTACCCCGGAACAGGTGGTCGCCATTGCAAGCCATGATGGAGGAAAGCAGGCGCTCGAAACAGTCCAGGCACTTTTGCCCGTACTTTGTCAAGCTCACGGTCTCACCCCGGAACAGGTGGTAGCCATTGCATCTAACATCGGAGGTAAGCAAGCATTGGAAACGGTTCAGGCCCTGTTGCCTGTACTTTGCCAGGCGCACGGTCTGACACCTGAGCAGGTTGTCGCCATCGCTAGCAACGGAGGTGGGAAACAGGCACTTGAAACTGTGCAGAGGCTTCTGCCGGTGCTGTGCCAAGCGCATGGCCTTACACCCGAGCAAGTAGTGGCTATTGCGAGTCATGATGGAGGCAAGCAAGCGCTGGAGACTGTCCAACGACTTCTTCCGGTCTTGTGTCAGGCACATGGATTGACCCCTCAACAAGTCGTGGCGATAGCTAGCAACGGCGGTGGAAAACAGGCCCTCGAAACCGTCCAGCGACTGCTCCCCGTACTGTGTCAAGCCCATGGACTTACCCCAGAACAAGTTGTGGCGATTGCCTCTAACAATGGTGGGAAGCAAGCTCTTGAGACGGTGCAGGCGTTGTTGCCCGTGCTTTGTCAAGCTCACGGGCTCACGCCAGAGCAAGTGGTCGCTATCGCGAGTAATAAAGGGGGCAAACAAGCCTTGGAGACAGTGCAAAGGCTCCTGCCAGTGCTCTGCCAGGCTCATGGTTTGACACCCGAACAGGTAGTTGCAATAGCGAGTCATGATGGCGGAAAGCAAGCTCTTGAAACTGTGCAGCGGCTGTTGCCTGTACTGTGTCAAGCCCACGGGCTGACACCGGAACAAGTTGTAGCGATCGCTAGCCACGATGGCGGGAAACAAGCTCTGGAAACGGTACAGAGACTCCTCCCAGTGCTTTGTCAGGCACACGGCCTCACGCCAGAGCAGGTTGTCGCCATCGCGTCAAACAATGGTGGAAAGCAGGCCCTGGAGACAGTCCAACGGTTGCTGCCGGTCCTTTGCCAGGCTCACGGGTTGACCCCCCAGCAGGTCGTGGCCATTGCCTCAAACAAGGGCGGTAGGCCAGCATTGGAGACGGTGCAGAGGCTTCTGCCTGTGCTCTGCCAAGCGCATGGACTCACCCCCGAGCAAGTGGTTGCTATCGCAAGTAACAACGGAGGGAAACAAGCGCTCGAAACCGTGCAAAGGTTGCTCCCCGTTCTCTGTCAGGCGCACGGTCTTACGCCACAACAGGTGGTGGCGATTGCATCTAATGGAGGCGGACGCCCTGCCTTGGAGAGCATTGTGGCCCAGCTGTCCAGGCCGGACCCTGCCCTGGCCGCGTTAACCAACGACCACCTCGTCGCCTTGGCCTGCCTCGGCGGAGGTTCTGGCGGCAGCGGATCCCAGCTGGTGAAGAGCGAGCTGGAGGAGAAGAAGTCCGAGCTGCGGCACAAGCTGAAGTACGTGCCCCACGAGTACATCGAGCTGATCGAGATCGCCAGGAACAGCACCCAGGACCGCATCCTGGAGATGAAGGTGATGGAGTTCTTCATGAAGGTGTACGGCTACAGGGGAAAGCACCTGGGCGGAAGCAGAAAGCCTGACGGCGCCATCTATACAGTGGGCAGCCCCATCGATTACGGCGTGATCGTGGACACAAAGGCCTACAGCGGCGGCTACAATCTGCCTATCGGCCAGGCCGACGAGATGCAGAGATACGTGGAGGAGAACCAGACCCGGAATAAGCACATCAACCCCAACGAGTGGTGGAAGGTGTACCCTAGCAGCGTGACCGAGTTCAAGTTCCTGTTCGTGAGCGGCCACTTCAAGGGCAACTACAAGGCCCAGCTGACCAGGCTGAACCACATCACCAACTGCAATGGCGCCGTGCTGAGCGTGGAGGAGCTGCTGATCGGCGGCGAGATGATCAAAGCCGGCACCCTGACACTGGAGGAGGTGCGGCGCAAGTTCAACAACGGCGAGATCAACTTCAGATCTTGATAACTCGAGTCTAGAGGGCCCGTTTAAACCCGCTGATCAGCCTCGACTGTGCCTTCTAGTTGCCAGCCATCTGTTGTTTGCCCCTCCCCCGTGCCTTCCTTGACCCTGGAAGGTGCCACTCCCACTGTCCTTTCCTAATAAAATGAGGAAATTGCATCGCATTGTCTGAGTAGGTGTCATTCTATTCTGGGGGGTGGGGTGGGGCAGGACAGCAAGGGGGAGGATTGGGAAGACAATAGCAGGCATGCTGGGGATGCGGTGGGCTCTATGGCTTCTACTGGGCGGTTTTATGGACAGCAAGCGAACCGGAATTGCCAGCTGGGGCGCCCTCTGGTAAGGTTGGGAAGCCCTGCAAAGTAAACTGGATGGCTTTCTCGCCGCCAAGGATCTGATGGCGCAGGGGATCAAGCTCTGATCAAGAGACAGGATGAGGATCGTTTCGCATGATTGAACAAGATGGATTGCACGCAGGTTCTCCGGCCGCTTGGGTGGAGAGGCTATTCGGCTATGACTGGGCACAACAGACAATCGGCTGCTCTGATGCCGCCGTGTTCCGGCTGTCAGCGCAGGGGCGCCCGGTTCTTTTTGTCAAGACCGACCTGTCCGGTGCCCTGAATGAACTGCAAGACGAGGCAGCGCGGCTATCGTGGCTGGCCACGACGGGCGTTCCTTGCGCAGCTGTGCTCGACGTTGTCACTGAAGCGGGAAGGGACTGGCTGCTATTGGGCGAAGTGCCGGGGCAGGATCTCCTGTCATCTCACCTTGCTCCTGCCGAGAAAGTATCCATCATGGCTGATGCAATGCGGCGGCTGCATACGCTTGATCCGGCTACCTGCCCATTCGACCACCAAGCGAAACATCGCATCGAGCGAGCACGTACTCGGATGGAAGCCGGTCTTGTCGATCAGGATGATCTGGACGAAGAGCATCAGGGGCTCGCGCCAGCCGAACTGTTCGCCAGGCTCAAGGCGAGCATGCCCGACGGCGAGGATCTCGTCGTGACCCATGGCGATGCCTGCTTGCCGAATATCATGGTGGAAAATGGCCGCTTTTCTGGATTCATCGACTGTGGCCGGCTGGGTGTGGCGGACCGCTATCAGGACATAGCGTTGGCTACCCGTGATATTGCTGAAGAGCTTGGCGGCGAATGGGCTGACCGCTTCCTCGTGCTTTACGGTATCGCCGCTCCCGATTCGCAGCGCATCGCCTTCTATCGCCTTCTTGACGAGTTCTTCTGAATTATTAACGCTTACAATTTCCTGATGCGGTATTTTCTCCTTACGCATCTGTGCGGTATTTCACACCGCATACAGGTGGCACTTTTCGGGGAAATGTGCGCGGAACCCCTATTTGTTTATTTTTCTAAATACATTCAAATATGTATCCGCTCATGAGACAATAACCCTGATAAATGCTTCAATAATAGCACGTGCTAAAACTTCATTTTTAATTTAAAAGGATCTAGGTGAAGATCCTTTTTGATAATCTCATGACCAAAATCCCTTAACGTGAGTTTTCGTTCCACTGAGCGTCAGACCCCGTAGAAAAGATCAAAGGATCTTCTTGAGATCCTTTTTTTCTGCGCGTAATCTGCTGCTTGCAAACAAAAAAACCACCGCTACCAGCGGTGGTTTGTTTGCCGGATCAAGAGCTACCAACTCTTTTTCCGAAGGTAACTGGCTTCAGCAGAGCGCAGATACCAAATACTGTCCTTCTAGTGTAGCCGTAGTTAGGCCACCACTTCAAGAACTCTGTAGCACCGCCTACATACCTCGCTCTGCTAATCCTGTTACCAGTGGCTGCTGCCAGTGGCGATAAGTCGTGTCTTACCGGGTTGGACTCAAGACGATAGTTACCGGATAAGGCGCAGCGGTCGGGCTGAACGGGGGGTTCGTGCACACAGCCCAGCTTGGAGCGAACGACCTACACCGAACTGAGATACCTACAGCGTGAGCTATGAGAAAGCGCCACGCTTCCCGAAGGGAGAAAGGCGGACAGGTATCCGGTAAGCGGCAGGGTCGGAACAGGAGAGCGCACGAGGGAGCTTCCAGGGGGAAACGCCTGGTATCTTTATAGTCCTGTCGGGTTTCGCCACCTCTGACTTGAGCGTCGATTTTTGTGATGCTCGTCAGGGGGGCGGAGCCTATGGAAAAACGCCAGCAACGCGGCCTTTTTACGGTTCCTGGGCTTTTGCTGGCCTTTTGCTCACATGTTCTT
それぞれの発現構築物の配列を再生成するために、上述の構築物の下線を引いた領域を以下に示されるそれぞれのCDSに置換する。
>NT L+28(配列番号218)
MDYKDHDGDYKDHDIDYKDDDDKMAPKKKRKVGIHGVPAAVDLRTLGYSQQQQEKIKPKVRSTVAQHHEALVGHGFTHAHIVALSQHPAALGTVAVKYQDMIAALPEATHEAIVGVGKQWSGARALEALLTVAGELRGPPLQLDTGQLLKIAKRGGVTAVEAVHAWRNALTGAPLNLTPEQVVAIASNNGGKQALETVQRLLPVLCQAHGLTPQQVVAIASNNGGKQALETVQRLLPVLCQAHGLTPQQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNNGGKQALETVQRLLPVLCQAHGLTPEQVVAIASHDGGKQALETVQALLPVLCQAHGLTPEQVVAIASHDGGKQALETVQALLPVLCQAHGLTPEQVVAIASNIGGKQALETVQALLPVLCQAHGLTPEQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPQQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNNGGKQALETVQALLPVLCQAHGLTPEQVVAIASNKGGKQALETVQRLLPVLCQAHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNNGGKQALETVQRLLPVLCQAHGLTPQQVVAIASNKGGRPALETVQRLLPVLCQAHGLTPEQVVAIASNNGGKQALETVQRLLPVLCQAHGLTPQQVVAIASNGGGRPALESIVAQLSRPDPALAALTNDHLVALACLGGGSGGSGSQLVKSELEEKKSELRHKLKYVPHEYIELIEIARNSTQDRILEMKVMEFFMKVYGYRGKHLGGSRKPDGAIYTVGSPIDYGVIVDTKAYSGGYNLPIGQADEMQRYVEENQTRNKHINPNEWWKVYPSSVTEFKFLFVSGHFKGNYKAQLTRLNHITNCNGAVLSVEELLIGGEMIKAGTLTLEEVRRKFNNGEINFRS
ATGGACTACAAAGACCATGACGGTGATTATAAAGATCATGACATCGATTACAAGGATGACGATGACAAGATGGCCCCCAAGAAGAAGAGGAAGGTGGGCATTCACGGGGTACCCGCCGCTGTGGATCTACGCACGCTCGGCTACAGCCAGCAGCAACAGGAGAAGATCAAACCGAAGGTTCGTTCGACAGTGGCGCAGCACCACGAGGCACTGGTCGGCCATGGGTTTACACACGCGCACATCGTTGCGCTCAGCCAACACCCGGCAGCGTTAGGGACCGTCGCTGTCAAGTATCAGGACATGATCGCAGCGTTGCCAGAGGCGACACACGAAGCGATCGTTGGCGTCGGCAAACAGTGGTCCGGCGCACGCGCCCTGGAGGCCTTGCTCACGGTGGCGGGAGAGTTGAGAGGTCCACCGTTACAGTTGGACACAGGCCAACTTCTCAAGATTGCAAAACGTGGCGGCGTGACCGCAGTGGAGGCAGTGCATGCATGGCGCAATGCACTGACGGGGGCCCCCCTGAACCTTACTCCCGAACAAGTAGTAGCGATAGCCAGTAATAACGGAGGTAAACAAGCCTTGGAGACGGTCCAAAGGTTGCTCCCGGTCTTGTGTCAGGCACATGGGCTGACGCCTCAACAGGTCGTCGCGATAGCGTCTAATAATGGAGGAAAGCAAGCTCTGGAAACCGTCCAGCGACTCCTTCCGGTTCTGTGCCAGGCTCATGGTCTGACTCCGCAGCAAGTCGTTGCTATAGCGTCCAACATCGGAGGCAAACAGGCCCTGGAGACCGTGCAGCGGTTGTTGCCTGTGCTTTGCCAAGCCCACGGGCTTACGCCTGAGCAAGTGGTGGCGATTGCCAGTAACAACGGCGGCAAACAAGCCCTTGAGACTGTGCAGAGGCTCTTGCCGGTACTCTGCCAAGCACACGGCTTGACCCCCGAGCAGGTTGTAGCCATAGCTAGTCACGACGGGGGTAAGCAAGCGTTGGAAACGGTGCAAGCACTTCTCCCCGTTCTCTGTCAAGCGCATGGACTTACCCCGGAACAGGTGGTCGCCATTGCAAGCCATGATGGAGGAAAGCAGGCGCTCGAAACAGTCCAGGCACTTTTGCCCGTACTTTGTCAAGCTCACGGTCTCACCCCGGAACAGGTGGTAGCCATTGCATCTAACATCGGAGGTAAGCAAGCATTGGAAACGGTTCAGGCCCTGTTGCCTGTACTTTGCCAGGCGCACGGTCTGACACCTGAGCAGGTTGTCGCCATCGCTAGCAACGGAGGTGGGAAACAGGCACTTGAAACTGTGCAGAGGCTTCTGCCGGTGCTGTGCCAAGCGCATGGCCTTACACCCGAGCAAGTAGTGGCTATTGCGAGTCATGATGGAGGCAAGCAAGCGCTGGAGACTGTCCAACGACTTCTTCCGGTCTTGTGTCAGGCACATGGATTGACCCCTCAACAAGTCGTGGCGATAGCTAGCAACGGCGGTGGAAAACAGGCCCTCGAAACCGTCCAGCGACTGCTCCCCGTACTGTGTCAAGCCCATGGACTTACCCCAGAACAAGTTGTGGCGATTGCCTCTAACAATGGTGGGAAGCAAGCTCTTGAGACGGTGCAGGCGTTGTTGCCCGTGCTTTGTCAAGCTCACGGGCTCACGCCAGAGCAAGTGGTCGCTATCGCGAGTAATAAAGGGGGCAAACAAGCCTTGGAGACAGTGCAAAGGCTCCTGCCAGTGCTCTGCCAGGCTCATGGTTTGACACCCGAACAGGTAGTTGCAATAGCGAGTCATGATGGCGGAAAGCAAGCTCTTGAAACTGTGCAGCGGCTGTTGCCTGTACTGTGTCAAGCCCACGGGCTGACACCGGAACAAGTTGTAGCGATCGCTAGCCACGATGGCGGGAAACAAGCTCTGGAAACGGTACAGAGACTCCTCCCAGTGCTTTGTCAGGCACACGGCCTCACGCCAGAGCAGGTTGTCGCCATCGCGTCAAACAATGGTGGAAAGCAGGCCCTGGAGACAGTCCAACGGTTGCTGCCGGTCCTTTGCCAGGCTCACGGGTTGACCCCCCAGCAGGTCGTGGCCATTGCCTCAAACAAGGGCGGTAGGCCAGCATTGGAGACGGTGCAGAGGCTTCTGCCTGTGCTCTGCCAAGCGCATGGACTCACCCCCGAGCAAGTGGTTGCTATCGCAAGTAACAACGGAGGGAAACAAGCGCTCGAAACCGTGCAAAGGTTGCTCCCCGTTCTCTGTCAGGCGCACGGTCTTACGCCACAACAGGTGGTGGCGATTGCATCTAATGGAGGCGGACGCCCTGCCTTGGAGAGCATTGTGGCCCAGCTGTCCAGGCCGGACCCTGCCCTGGCCGCGTTAACCAACGACCACCTCGTCGCCTTGGCCTGCCTCGGCGGAGGTTCTGGCGGCAGCGGATCCCAGCTGGTGAAGAGCGAGCTGGAGGAGAAGAAGTCCGAGCTGCGGCACAAGCTGAAGTACGTGCCCCACGAGTACATCGAGCTGATCGAGATCGCCAGGAACAGCACCCAGGACCGCATCCTGGAGATGAAGGTGATGGAGTTCTTCATGAAGGTGTACGGCTACAGGGGAAAGCACCTGGGCGGAAGCAGAAAGCCTGACGGCGCCATCTATACAGTGGGCAGCCCCATCGATTACGGCGTGATCGTGGACACAAAGGCCTACAGCGGCGGCTACAATCTGCCTATCGGCCAGGCCGACGAGATGCAGAGATACGTGGAGGAGAACCAGACCCGGAATAAGCACATCAACCCCAACGAGTGGTGGAAGGTGTACCCTAGCAGCGTGACCGAGTTCAAGTTCCTGTTCGTGAGCGGCCACTTCAAGGGCAACTACAAGGCCCAGCTGACCAGGCTGAACCACATCACCAACTGCAATGGCGCCGTGCTGAGCGTGGAGGAGCTGCTGATCGGCGGCGAGATGATCAAAGCCGGCACCCTGACACTGGAGGAGGTGCGGCGCAAGTTCAACAACGGCGAGATCAACTTCAGATCT
MDYKDHDGDYKDHDIDYKDDDDKMAPKKKRKVGIHGVPMVDLRTLGYSQQQQEKIKPKVRSTVAQHHEALVGHGFTHAHIVALSQHPAALGTVAVKYQDMIAALPEATHEAIVGVGKQWSGARALEALLTVAGELRGPPLQLDTGQLLKIAKRGGVTAVEAVHAWRNALTGAPLNLTPEQVVAIASNNGGKQALETVQRLLPVLCQAHGLTPQQVVAIASNNGGKQALETVQRLLPVLCQAHGLTPQQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNNGGKQALETVQRLLPVLCQAHGLTPEQVVAIASHDGGKQALETVQALLPVLCQAHGLTPEQVVAIASHDGGKQALETVQALLPVLCQAHGLTPEQVVAIASNIGGKQALETVQALLPVLCQAHGLTPEQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPQQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNNGGKQALETVQALLPVLCQAHGLTPEQVVAIASNKGGKQALETVQRLLPVLCQAHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNNGGKQALETVQRLLPVLCQAHGLTPQQVVAIASNKGGRPALETVQRLLPVLCQAHGLTPEQVVAIASNNGGKQALETVQRLLPVLCQAHGLTPQQVVAIASNGGGRPALESIVAQLSRPDPALAALTNDHLVALACLGGRPALDAVKKGLPHAPALIKRTNRRIPERTSHRVAGSQLVKSELEEKKSELRHKLKYVPHEYIELIEIARNSTQDRILEMKVMEFFMKVYGYRGKHLGGSRKPDGAIYTVGSPIDYGVIVDTKAYSGGYNLPIGQADEMQRYVEENQTRNKHINPNEWWKVYPSSVTEFKFLFVSGHFKGNYKAQLTRLNHITNCNGAVLSVEELLIGGEMIKAGTLTLEEVRRKFNNGEINFRS
ATGGACTACAAAGACCATGACGGTGATTATAAAGATCATGACATCGATTACAAGGATGACGATGACAAGATGGCCCCCAAGAAGAAGAGGAAGGTGGGCATTCACGGGGTACCCATGGTGGATCTACGCACGCTCGGCTACAGCCAGCAGCAACAGGAGAAGATCAAACCGAAGGTTCGTTCGACAGTGGCGCAGCACCACGAGGCACTGGTCGGCCATGGGTTTACACACGCGCACATCGTTGCGCTCAGCCAACACCCGGCAGCGTTAGGGACCGTCGCTGTCAAGTATCAGGACATGATCGCAGCGTTGCCAGAGGCGACACACGAAGCGATCGTTGGCGTCGGCAAACAGTGGTCCGGCGCACGCGCCCTGGAGGCCTTGCTCACGGTGGCGGGAGAGTTGAGAGGTCCACCGTTACAGTTGGACACAGGCCAACTTCTCAAGATTGCAAAACGTGGCGGCGTGACCGCAGTGGAGGCAGTGCATGCATGGCGCAATGCACTGACGGGGGCCCCCCTGAACCTTACTCCCGAACAAGTAGTAGCGATAGCCAGTAATAACGGAGGTAAACAAGCCTTGGAGACGGTCCAAAGGTTGCTCCCGGTCTTGTGTCAGGCACATGGGCTGACGCCTCAACAGGTCGTCGCGATAGCGTCTAATAATGGAGGAAAGCAAGCTCTGGAAACCGTCCAGCGACTCCTTCCGGTTCTGTGCCAGGCTCATGGTCTGACTCCGCAGCAAGTCGTTGCTATAGCGTCCAACATCGGAGGCAAACAGGCCCTGGAGACCGTGCAGCGGTTGTTGCCTGTGCTTTGCCAAGCCCACGGGCTTACGCCTGAGCAAGTGGTGGCGATTGCCAGTAACAACGGCGGCAAACAAGCCCTTGAGACTGTGCAGAGGCTCTTGCCGGTACTCTGCCAAGCACACGGCTTGACCCCCGAGCAGGTTGTAGCCATAGCTAGTCACGACGGGGGTAAGCAAGCGTTGGAAACGGTGCAAGCACTTCTCCCCGTTCTCTGTCAAGCGCATGGACTTACCCCGGAACAGGTGGTCGCCATTGCAAGCCATGATGGAGGAAAGCAGGCGCTCGAAACAGTCCAGGCACTTTTGCCCGTACTTTGTCAAGCTCACGGTCTCACCCCGGAACAGGTGGTAGCCATTGCATCTAACATCGGAGGTAAGCAAGCATTGGAAACGGTTCAGGCCCTGTTGCCTGTACTTTGCCAGGCGCACGGTCTGACACCTGAGCAGGTTGTCGCCATCGCTAGCAACGGAGGTGGGAAACAGGCACTTGAAACTGTGCAGAGGCTTCTGCCGGTGCTGTGCCAAGCGCATGGCCTTACACCCGAGCAAGTAGTGGCTATTGCGAGTCATGATGGAGGCAAGCAAGCGCTGGAGACTGTCCAACGACTTCTTCCGGTCTTGTGTCAGGCACATGGATTGACCCCTCAACAAGTCGTGGCGATAGCTAGCAACGGCGGTGGAAAACAGGCCCTCGAAACCGTCCAGCGACTGCTCCCCGTACTGTGTCAAGCCCATGGACTTACCCCAGAACAAGTTGTGGCGATTGCCTCTAACAATGGTGGGAAGCAAGCTCTTGAGACGGTGCAGGCGTTGTTGCCCGTGCTTTGTCAAGCTCACGGGCTCACGCCAGAGCAAGTGGTCGCTATCGCGAGTAATAAAGGGGGCAAACAAGCCTTGGAGACAGTGCAAAGGCTCCTGCCAGTGCTCTGCCAGGCTCATGGTTTGACACCCGAACAGGTAGTTGCAATAGCGAGTCATGATGGCGGAAAGCAAGCTCTTGAAACTGTGCAGCGGCTGTTGCCTGTACTGTGTCAAGCCCACGGGCTGACACCGGAACAAGTTGTAGCGATCGCTAGCCACGATGGCGGGAAACAAGCTCTGGAAACGGTACAGAGACTCCTCCCAGTGCTTTGTCAGGCACACGGCCTCACGCCAGAGCAGGTTGTCGCCATCGCGTCAAACAATGGTGGAAAGCAGGCCCTGGAGACAGTCCAACGGTTGCTGCCGGTCCTTTGCCAGGCTCACGGGTTGACCCCCCAGCAGGTCGTGGCCATTGCCTCAAACAAGGGCGGTAGGCCAGCATTGGAGACGGTGCAGAGGCTTCTGCCTGTGCTCTGCCAAGCGCATGGACTCACCCCCGAGCAAGTGGTTGCTATCGCAAGTAACAACGGAGGGAAACAAGCGCTCGAAACCGTGCAAAGGTTGCTCCCCGTTCTCTGTCAGGCGCACGGTCTTACGCCACAACAGGTGGTGGCGATTGCATCTAATGGAGGCGGACGCCCTGCCTTGGAGAGCATTGTGGCCCAGCTGTCCAGGCCGGACCCTGCCCTGGCCGCGTTAACCAACGACCACCTCGTCGCCTTGGCCTGCCTCGGCGGACGTCCTGCGCTGGATGCAGTGAAAAAGGGATTGCCGCACGCGCCGGCCTTGATCAAAAGAACCAATCGCCGTATTCCCGAACGCACATCCCATCGCGTTGCCGGATCCCAGCTGGTGAAGAGCGAGCTGGAGGAGAAGAAGTCCGAGCTGCGGCACAAGCTGAAGTACGTGCCCCACGAGTACATCGAGCTGATCGAGATCGCCAGGAACAGCACCCAGGACCGCATCCTGGAGATGAAGGTGATGGAGTTCTTCATGAAGGTGTACGGCTACAGGGGAAAGCACCTGGGCGGAAGCAGAAAGCCTGACGGCGCCATCTATACAGTGGGCAGCCCCATCGATTACGGCGTGATCGTGGACACAAAGGCCTACAGCGGCGGCTACAATCTGCCTATCGGCCAGGCCGACGAGATGCAGAGATACGTGGAGGAGAACCAGACCCGGAATAAGCACATCAACCCCAACGAGTGGTGGAAGGTGTACCCTAGCAGCGTGACCGAGTTCAAGTTCCTGTTCGTGAGCGGCCACTTCAAGGGCAACTACAAGGCCCAGCTGACCAGGCTGAACCACATCACCAACTGCAATGGCGCCGTGCTGAGCGTGGAGGAGCTGCTGATCGGCGGCGAGATGATCAAAGCCGGCACCCTGACACTGGAGGAGGTGCGGCGCAAGTTCAACAACGGCGAGATCAACTTCAGATCT
MDYKDHDGDYKDHDIDYKDDDDKMAPKKKRKVGIHGVPMVDLRTLGYSQQQQEKIKPKVRSTVAQHHEALVGHGFTHAHIVALSQHPAALGTVAVKYQDMIAALPEATHEAIVGVGKQWSGARALEALLTVAGELRGPPLQLDTGQLLKIAKRGGVTAVEAVHAWRNALTGAPLNLTPEQVVAIASNNGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPAQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNKGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPQQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNKGGKQALETVQRLLPVLCQAHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNNGGKQALETVQRLLPVLCQAHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPQQVVAIASNKGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNNGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNGGGRPALESIVAQLSRPDPALAALTNDHLVALACLGGSQLVKSELEEKKSELRHKLKYVPHEYIELIEIARNSTQDRILEMKVMEFFMKVYGYRGKHLGGSRKPDGAIYTVGSPIDYGVIVDTKAYSGGYNLPIGQADEMQRYVEENQTRNKHINPNEWWKVYPSSVTEFKFLFVSGHFKGNYKAQLTRLNHITNCNGAVLSVEELLIGGEMIKAGTLTLEEVRRKFNNGEINFRS
ATGGACTACAAAGACCATGACGGTGATTATAAAGATCATGACATCGATTACAAGGATGACGATGACAAGATGGCCCCCAAGAAGAAGAGGAAGGTGGGCATTCACGGGGTACCTATGGTGGACTTGAGGACACTCGGTTATTCGCAACAGCAACAGGAGAAAATCAAGCCTAAGGTCAGGAGCACCGTCGCGCAACACCACGAGGCGCTTGTGGGGCATGGCTTCACTCATGCGCATATTGTCGCGCTTTCACAGCACCCTGCGGCGCTTGGGACGGTGGCTGTCAAATACCAAGATATGATTGCGGCCCTGCCCGAAGCCACGCACGAGGCAATTGTAGGGGTCGGTAAACAGTGGTCGGGAGCGCGAGCACTTGAGGCGCTGCTGACTGTGGCGGGTGAGCTTAGGGGGCCTCCGCTCCAGCTCGACACCGGGCAGCTGCTGAAGATCGCGAAGAGAGGGGGAGTAACAGCGGTAGAGGCAGTGCACGCCTGGCGCAATGCGCTCACCGGGGCCCCCTTGAATCTTACTCCAGAGCAGGTCGTCGCAATCGCGTCGAATAACGGGGGAAAGCAAGCACTGGAAACCGTGCAGAGGTTGTTGCCGGTCTTGTGTCAGGCTCACGGCTTGACACCTGCCCAAGTGGTGGCCATTGCGTCGAACATCGGGGGAAAACAGGCACTTGAAACAGTCCAGAGACTTTTGCCCGTCCTCTGCCAGGCGCACGGCCTCACGCCGGATCAGGTGGTAGCCATCGCGTCAAACATCGGAGGGAAGCAGGCTCTGGAAACGGTGCAGCGGCTTTTGCCGGTACTTTGCCAAGCTCATGGGCTCACGCCAGCCCAAGTGGTAGCTATCGCATCGCACGACGGAGGGAAGCAGGCCTTGGAGACAGTGCAACGGCTCCTCCCCGTGTTGTGCCAGGCACATGGGTTGACTCCAGAGCAGGTCGTAGCAATCGCCTCCAATATCGGGGGAAAGCAAGCGTTGGAGACAGTGCAGCGACTGCTGCCTGTGCTTTGCCAGGCTCATGGCCTGACGCCCGATCAGGTAGTGGCAATCGCGTCAAACAAAGGTGGAAAGCAGGCACTCGAAACGGTACAGCGCTTGCTGCCCGTCTTGTGTCAGGCCCACGGTCTGACACCCGACCAGGTAGTCGCGATTGCGTCGAACATCGGGGGAAAGCAAGCGTTGGAAACGGTACAACGCCTGCTCCCGGTGCTCTGCCAGGCTCATGGACTTACACCCGAGCAGGTGGTCGCCATCGCGTCAAACATCGGAGGCAAACAGGCATTGGAGACAGTGCAGCGCCTTCTCCCAGTCTTGTGTCAGGCCCACGGTCTGACACCCGACCAGGTCGTCGCGATTGCATCGAATGGAGGTGGGAAACAGGCCCTTGAGACAGTACAGAGGCTTTTGCCCGTGTTGTGCCAGGCCCACGGACTCACACCCGAACAAGTCGTCGCCATTGCCAGCCATGATGGAGGTAAACAGGCACTTGAGACTGTCCAGCGCCTCCTGCCGGTGCTGTGCCAAGCACATGGGCTGACCCCGCAGCAAGTCGTAGCGATCGCCTCGAATGGTGGAGGAAAACAAGCGCTTGAAACCGTCCAGAGGTTGCTCCCGGTGCTGTGCCAGGCACATGGCCTTACGCCTGAACAAGTAGTCGCGATTGCCAGCAACAAAGGCGGAAAACAGGCTCTCGAAACGGTCCAGCGGTTGCTGCCGGTGTTGTGCCAGGCGCACGGTCTTACACCGGACCAGGTGGTGGCGATTGCCTCCCACGATGGGGGTAAACAGGCACTGGAAACCGTGCAGAGATTGCTCCCAGTACTTTGTCAGGCACATGGTCTGACTCCTGCTCAAGTGGTCGCGATCGCCTCGAACAATGGCGGAAAGCAGGCGCTCGAAACGGTACAGCGGCTCCTTCCGGTGCTCTGCCAAGCCCACGGATTGACGCCAGAACAGGTCGTGGCAATTGCGTCACACGACGGTGGAAAGCAGGCGCTCGAAACTGTGCAAAGACTCCTGCCCGTACTCTGCCAGGCACACGGTTTGACTCCCCAGCAGGTAGTGGCCATCGCGAGCAATAAGGGAGGAAAGCAGGCGCTTGAAACGGTGCAGAGACTTCTGCCCGTGCTTTGTCAAGCCCACGGGCTGACTCCGGAGCAGGTAGTGGCCATCGCCTCAAACAACGGAGGAAAGCAAGCTCTCGAAACCGTACAGAGGCTTCTCCCCGTGCTCTGTCAGGCCCACGGGTTGACCCCCGAACAGGTGGTCGCCATTGCTAGCAACGGGGGAGGACGGCCAGCCTTGGAGTCCATCGTAGCCCAATTGTCCAGGCCCGATCCCGCGTTGGCTGCGTTAACGAATGACCATCTGGTGGCGTTGGCATGTCTTGGTGGATCCCAGCTGGTGAAGAGCGAGCTGGAGGAGAAGAAGTCCGAGCTGCGGCACAAGCTGAAGTACGTGCCCCACGAGTACATCGAGCTGATCGAGATCGCCAGGAACAGCACCCAGGACCGCATCCTGGAGATGAAGGTGATGGAGTTCTTCATGAAGGTGTACGGCTACAGGGGAAAGCACCTGGGCGGAAGCAGAAAGCCTGACGGCGCCATCTATACAGTGGGCAGCCCCATCGATTACGGCGTGATCGTGGACACAAAGGCCTACAGCGGCGGCTACAATCTGCCTATCGGCCAGGCCGACGAGATGCAGAGATACGTGGAGGAGAACCAGACCCGGAATAAGCACATCAACCCCAACGAGTGGTGGAAGGTGTACCCTAGCAGCGTGACCGAGTTCAAGTTCCTGTTCGTGAGCGGCCACTTCAAGGGCAACTACAAGGCCCAGCTGACCAGGCTGAACCACATCACCAACTGCAATGGCGCCGTGCTGAGCGTGGAGGAGCTGCTGATCGGCGGCGAGATGATCAAAGCCGGCACCCTGACACTGGAGGAGGTGCGGCGCAAGTTCAACAACGGCGAGATCAACTTCAGATCT
MDYKDHDGDYKDHDIDYKDDDDKMAPKKKRKVGIHGVPMVDLRTLGYSQQQQEKIKPKVRSTVAQHHEALVGHGFTHAHIVALSQHPAALGTVAVKYQDMIAALPEATHEAIVGVGKQWSGARALEALLTVAGELRGPPLQLDTGQLLKIAKRGGVTAVEAVHAWRNALTGAPLNLTPEQVVAIASNNGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPAQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNKGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPQQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNKGGKQALETVQRLLPVLCQAHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNNGGKQALETVQRLLPVLCQAHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPQQVVAIASNKGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNNGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNGGGRPALESIVAQLSRPDPALAALTNDHLVALACLGGRPALDAVKKGLPHAPALIKRTNRRIPERTSHRVAGSQLVKSELEEKKSELRHKLKYVPHEYIELIEIARNSTQDRILEMKVMEFFMKVYGYRGKHLGGSRKPDGAIYTVGSPIDYGVIVDTKAYSGGYNLPIGQADEMQRYVEENQTRNKHINPNEWWKVYPSSVTEFKFLFVSGHFKGNYKAQLTRLNHITNCNGAVLSVEELLIGGEMIKAGTLTLEEVRRKFNNGEINFRS
ATGGACTACAAAGACCATGACGGTGATTATAAAGATCATGACATCGATTACAAGGATGACGATGACAAGATGGCCCCCAAGAAGAAGAGGAAGGTGGGCATTCACGGGGTACCTATGGTGGACTTGAGGACACTCGGTTATTCGCAACAGCAACAGGAGAAAATCAAGCCTAAGGTCAGGAGCACCGTCGCGCAACACCACGAGGCGCTTGTGGGGCATGGCTTCACTCATGCGCATATTGTCGCGCTTTCACAGCACCCTGCGGCGCTTGGGACGGTGGCTGTCAAATACCAAGATATGATTGCGGCCCTGCCCGAAGCCACGCACGAGGCAATTGTAGGGGTCGGTAAACAGTGGTCGGGAGCGCGAGCACTTGAGGCGCTGCTGACTGTGGCGGGTGAGCTTAGGGGGCCTCCGCTCCAGCTCGACACCGGGCAGCTGCTGAAGATCGCGAAGAGAGGGGGAGTAACAGCGGTAGAGGCAGTGCACGCCTGGCGCAATGCGCTCACCGGGGCCCCCTTGAATCTTACTCCAGAGCAGGTCGTCGCAATCGCGTCGAATAACGGGGGAAAGCAAGCACTGGAAACCGTGCAGAGGTTGTTGCCGGTCTTGTGTCAGGCTCACGGCTTGACACCTGCCCAAGTGGTGGCCATTGCGTCGAACATCGGGGGAAAACAGGCACTTGAAACAGTCCAGAGACTTTTGCCCGTCCTCTGCCAGGCGCACGGCCTCACGCCGGATCAGGTGGTAGCCATCGCGTCAAACATCGGAGGGAAGCAGGCTCTGGAAACGGTGCAGCGGCTTTTGCCGGTACTTTGCCAAGCTCATGGGCTCACGCCAGCCCAAGTGGTAGCTATCGCATCGCACGACGGAGGGAAGCAGGCCTTGGAGACAGTGCAACGGCTCCTCCCCGTGTTGTGCCAGGCACATGGGTTGACTCCAGAGCAGGTCGTAGCAATCGCCTCCAATATCGGGGGAAAGCAAGCGTTGGAGACAGTGCAGCGACTGCTGCCTGTGCTTTGCCAGGCTCATGGCCTGACGCCCGATCAGGTAGTGGCAATCGCGTCAAACAAAGGTGGAAAGCAGGCACTCGAAACGGTACAGCGCTTGCTGCCCGTCTTGTGTCAGGCCCACGGTCTGACACCCGACCAGGTAGTCGCGATTGCGTCGAACATCGGGGGAAAGCAAGCGTTGGAAACGGTACAACGCCTGCTCCCGGTGCTCTGCCAGGCTCATGGACTTACACCCGAGCAGGTGGTCGCCATCGCGTCAAACATCGGAGGCAAACAGGCATTGGAGACAGTGCAGCGCCTTCTCCCAGTCTTGTGTCAGGCCCACGGTCTGACACCCGACCAGGTCGTCGCGATTGCATCGAATGGAGGTGGGAAACAGGCCCTTGAGACAGTACAGAGGCTTTTGCCCGTGTTGTGCCAGGCCCACGGACTCACACCCGAACAAGTCGTCGCCATTGCCAGCCATGATGGAGGTAAACAGGCACTTGAGACTGTCCAGCGCCTCCTGCCGGTGCTGTGCCAAGCACATGGGCTGACCCCGCAGCAAGTCGTAGCGATCGCCTCGAATGGTGGAGGAAAACAAGCGCTTGAAACCGTCCAGAGGTTGCTCCCGGTGCTGTGCCAGGCACATGGCCTTACGCCTGAACAAGTAGTCGCGATTGCCAGCAACAAAGGCGGAAAACAGGCTCTCGAAACGGTCCAGCGGTTGCTGCCGGTGTTGTGCCAGGCGCACGGTCTTACACCGGACCAGGTGGTGGCGATTGCCTCCCACGATGGGGGTAAACAGGCACTGGAAACCGTGCAGAGATTGCTCCCAGTACTTTGTCAGGCACATGGTCTGACTCCTGCTCAAGTGGTCGCGATCGCCTCGAACAATGGCGGAAAGCAGGCGCTCGAAACGGTACAGCGGCTCCTTCCGGTGCTCTGCCAAGCCCACGGATTGACGCCAGAACAGGTCGTGGCAATTGCGTCACACGACGGTGGAAAGCAGGCGCTCGAAACTGTGCAAAGACTCCTGCCCGTACTCTGCCAGGCACACGGTTTGACTCCCCAGCAGGTAGTGGCCATCGCGAGCAATAAGGGAGGAAAGCAGGCGCTTGAAACGGTGCAGAGACTTCTGCCCGTGCTTTGTCAAGCCCACGGGCTGACTCCGGAGCAGGTAGTGGCCATCGCCTCAAACAACGGAGGAAAGCAAGCTCTCGAAACCGTACAGAGGCTTCTCCCCGTGCTCTGTCAGGCCCACGGGTTGACCCCCGAACAGGTGGTCGCCATTGCTAGCAACGGGGGAGGACGGCCAGCCTTGGAGTCCATCGTAGCCCAATTGTCCAGGCCCGATCCCGCGTTGGCTGCGTTAACGAATGACCATCTGGTGGCGTTGGCATGTCTTGGTGGACGACCCGCGCTCGATGCAGTCAAAAAGGGTCTGCCTCATGCTCCCGCATTGATCAAAAGAACCAACCGGCGGATTCCCGAGAGAACTTCCCATCGAGTCGCGGGATCCCAGCTGGTGAAGAGCGAGCTGGAGGAGAAGAAGTCCGAGCTGCGGCACAAGCTGAAGTACGTGCCCCACGAGTACATCGAGCTGATCGAGATCGCCAGGAACAGCACCCAGGACCGCATCCTGGAGATGAAGGTGATGGAGTTCTTCATGAAGGTGTACGGCTACAGGGGAAAGCACCTGGGCGGAAGCAGAAAGCCTGACGGCGCCATCTATACAGTGGGCAGCCCCATCGATTACGGCGTGATCGTGGACACAAAGGCCTACAGCGGCGGCTACAATCTGCCTATCGGCCAGGCCGACGAGATGCAGAGATACGTGGAGGAGAACCAGACCCGGAATAAGCACATCAACCCCAACGAGTGGTGGAAGGTGTACCCTAGCAGCGTGACCGAGTTCAAGTTCCTGTTCGTGAGCGGCCACTTCAAGGGCAACTACAAGGCCCAGCTGACCAGGCTGAACCACATCACCAACTGCAATGGCGCCGTGCTGAGCGTGGAGGAGCTGCTGATCGGCGGCGAGATGATCAAAGCCGGCACCCTGACACTGGAGGAGGTGCGGCGCAAGTTCAACAACGGCGAGATCAACTTCAGATCT
MDYKDHDGDYKDHDIDYKDDDDKMAPKKKRKVGIHGVPMVDLRTLGYSQQQQEKIKPKVRSTVAQHHEALVGHGFTHAHIVALSQHPAALGTVAVKYQDMIAALPEATHEAIVGVGKQWSGARALEALLTVAGELRGPPLQLDTGQLLKIAKRGGVTAVEAVHAWRNALTGAPLNLTPEQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNIGGKQALETVQRLLPVLRQAHGLTPEQVVAIASNGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNGGGRPALESIVAQLSRPDPSLAALTNDHLVALACLGGSQLVKSELEEKKSELRHKLKYVPHEYIELIEIARNSTQDRILEMKVMEFFMKVYGYRGKHLGGSRKPDGAIYTVGSPIDYGVIVDTKAYSGGYNLPIGQADEMQRYVEENQTRNKHINPNEWWKVYPSSVTEFKFLFVSGHFKGNYKAQLTRLNHITNCNGAVLSVEELLIGGEMIKAGTLTLEEVRRKFNNGEINFRS
ATGGACTACAAAGACCATGACGGTGATTATAAAGATCATGACATCGATTACAAGGATGACGATGACAAGATGGCCCCCAAGAAGAAGAGGAAGGTGGGCATTCACGGGGTACCCATGGTGGATCTACGCACGCTCGGCTACAGCCAGCAGCAACAGGAGAAGATCAAACCGAAGGTTCGTTCGACAGTGGCGCAGCACCACGAGGCACTGGTCGGCCATGGGTTTACACACGCGCACATCGTTGCGCTCAGCCAACACCCGGCAGCGTTAGGGACCGTCGCTGTCAAGTATCAGGACATGATCGCAGCGTTGCCAGAGGCGACACACGAAGCGATCGTTGGCGTCGGCAAACAGTGGTCCGGCGCACGCGCCCTGGAGGCCTTGCTCACGGTGGCGGGAGAGTTGAGAGGTCCACCGTTACAGTTGGACACAGGCCAACTTCTCAAGATTGCAAAACGTGGCGGCGTGACCGCAGTGGAGGCAGTGCATGCATGGCGCAATGCACTGACGGGGGCCCCCCTGAACCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCAATATTGGTGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCAATGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCAATATTGGTGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCAATATTGGTGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGCACAGGTGGTGGCCATCGCCAGCAATATTGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGCGCCAGGCCCATGGCCTGACCCCGGAGCAGGTCGTGGCCATCGCCAGCAATGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGCACAGGTGGTGGCCATCGCCAGCAATATTGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTCGTGGCCATCGCCAGCCACGATGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCCACGATGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGACCAGGTGGTGGCCATCGCCAGCAATGGCGGTGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGACCAGGTGGTGGCCATCGCCAGCAATGGCGGTGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCCACGATGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCAATGGCGGCGGCAGGCCGGCGCTGGAGAGCATTGTTGCCCAGTTATCTCGCCCTGATCCGTCGTTGGCCGCGTTAACCAACGACCACCTCGTCGCCTTGGCCTGCCTCGGCGGATCCCAGCTGGTGAAGAGCGAGCTGGAGGAGAAGAAGTCCGAGCTGCGGCACAAGCTGAAGTACGTGCCCCACGAGTACATCGAGCTGATCGAGATCGCCAGGAACAGCACCCAGGACCGCATCCTGGAGATGAAGGTGATGGAGTTCTTCATGAAGGTGTACGGCTACAGGGGAAAGCACCTGGGCGGAAGCAGAAAGCCTGACGGCGCCATCTATACAGTGGGCAGCCCCATCGATTACGGCGTGATCGTGGACACAAAGGCCTACAGCGGCGGCTACAATCTGCCTATCGGCCAGGCCGACGAGATGCAGAGATACGTGGAGGAGAACCAGACCCGGAATAAGCACATCAACCCCAACGAGTGGTGGAAGGTGTACCCTAGCAGCGTGACCGAGTTCAAGTTCCTGTTCGTGAGCGGCCACTTCAAGGGCAACTACAAGGCCCAGCTGACCAGGCTGAACCACATCACCAACTGCAATGGCGCCGTGCTGAGCGTGGAGGAGCTGCTGATCGGCGGCGAGATGATCAAAGCCGGCACCCTGACACTGGAGGAGGTGCGGCGCAAGTTCAACAACGGCGAGATCAACTTCAGATCT
MDYKDHDGDYKDHDIDYKDDDDKMAPKKKRKVGIHGVPMVDLRTLGYSQQQQEKIKPKVRSTVAQHHEALVGHGFTHAHIVALSQHPAALGTVAVKYQDMIAALPEATHEAIVGVGKQWSGARALEALLTVAGELRGPPLQLDTGQLLKIAKRGGVTAVEAVHAWRNALTGAPLNLTPEQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNIGGKQALETVQRLLPVLRQAHGLTPEQVVAIASNGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNGGGRPALESIVAQLSRPDPSLAALTNDHLVALACLGGRPALDAVKKGGSQLVKSELEEKKSELRHKLKYVPHEYIELIEIARNSTQDRILEMKVMEFFMKVYGYRGKHLGGSRKPDGAIYTVGSPIDYGVIVDTKAYSGGYNLPIGQADEMQRYVEENQTRNKHINPNEWWKVYPSSVTEFKFLFVSGHFKGNYKAQLTRLNHITNCNGAVLSVEELLIGGEMIKAGTLTLEEVRRKFNNGEINFRS
ATGGACTACAAAGACCATGACGGTGATTATAAAGATCATGACATCGATTACAAGGATGACGATGACAAGATGGCCCCCAAGAAGAAGAGGAAGGTGGGCATTCACGGGGTACCCATGGTGGATCTACGCACGCTCGGCTACAGCCAGCAGCAACAGGAGAAGATCAAACCGAAGGTTCGTTCGACAGTGGCGCAGCACCACGAGGCACTGGTCGGCCATGGGTTTACACACGCGCACATCGTTGCGCTCAGCCAACACCCGGCAGCGTTAGGGACCGTCGCTGTCAAGTATCAGGACATGATCGCAGCGTTGCCAGAGGCGACACACGAAGCGATCGTTGGCGTCGGCAAACAGTGGTCCGGCGCACGCGCCCTGGAGGCCTTGCTCACGGTGGCGGGAGAGTTGAGAGGTCCACCGTTACAGTTGGACACAGGCCAACTTCTCAAGATTGCAAAACGTGGCGGCGTGACCGCAGTGGAGGCAGTGCATGCATGGCGCAATGCACTGACGGGGGCCCCCCTGAACCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCAATATTGGTGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCAATGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCAATATTGGTGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCAATATTGGTGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGCACAGGTGGTGGCCATCGCCAGCAATATTGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGCGCCAGGCCCATGGCCTGACCCCGGAGCAGGTCGTGGCCATCGCCAGCAATGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGCACAGGTGGTGGCCATCGCCAGCAATATTGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTCGTGGCCATCGCCAGCCACGATGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCCACGATGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGACCAGGTGGTGGCCATCGCCAGCAATGGCGGTGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGACCAGGTGGTGGCCATCGCCAGCAATGGCGGTGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCCACGATGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCAATGGCGGCGGCAGGCCGGCGCTGGAGAGCATTGTTGCCCAGTTATCTCGCCCTGATCCGTCGTTGGCCGCGTTAACCAACGACCACCTCGTCGCCTTGGCCTGCCTCGGCGGACGTCCTGCGCTGGATGCAGTGAAAAAGGGAGGATCCCAGCTGGTGAAGAGCGAGCTGGAGGAGAAGAAGTCCGAGCTGCGGCACAAGCTGAAGTACGTGCCCCACGAGTACATCGAGCTGATCGAGATCGCCAGGAACAGCACCCAGGACCGCATCCTGGAGATGAAGGTGATGGAGTTCTTCATGAAGGTGTACGGCTACAGGGGAAAGCACCTGGGCGGAAGCAGAAAGCCTGACGGCGCCATCTATACAGTGGGCAGCCCCATCGATTACGGCGTGATCGTGGACACAAAGGCCTACAGCGGCGGCTACAATCTGCCTATCGGCCAGGCCGACGAGATGCAGAGATACGTGGAGGAGAACCAGACCCGGAATAAGCACATCAACCCCAACGAGTGGTGGAAGGTGTACCCTAGCAGCGTGACCGAGTTCAAGTTCCTGTTCGTGAGCGGCCACTTCAAGGGCAACTACAAGGCCCAGCTGACCAGGCTGAACCACATCACCAACTGCAATGGCGCCGTGCTGAGCGTGGAGGAGCTGCTGATCGGCGGCGAGATGATCAAAGCCGGCACCCTGACACTGGAGGAGGTGCGGCGCAAGTTCAACAACGGCGAGATCAACTTCAGATCT
MDYKDHDGDYKDHDIDYKDDDDKMAPKKKRKVGIHGVPMVDLRTLGYSQQQQEKIKPKVRSTVAQHHEALVGHGFTHAHIVALSQHPAALGTVAVKYQDMIAALPEATHEAIVGVGKQWSGARALEALLTVAGELRGPPLQLDTGQLLKIAKRGGVTAVEAVHAWRNALTGAPLNLTPEQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNIGGKQALETVQRLLPVLRQAHGLTPEQVVAIASNGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNGGGRPALESIVAQLSRPDPSLAALTNDHLVALACLGGRPALDAVKKGLPHAPALIKRTGSQLVKSELEEKKSELRHKLKYVPHEYIELIEIARNSTQDRILEMKVMEFFMKVYGYRGKHLGGSRKPDGAIYTVGSPIDYGVIVDTKAYSGGYNLPIGQADEMQRYVEENQTRNKHINPNEWWKVYPSSVTEFKFLFVSGHFKGNYKAQLTRLNHITNCNGAVLSVEELLIGGEMIKAGTLTLEEVRRKFNNGEINFRS
ATGGACTACAAAGACCATGACGGTGATTATAAAGATCATGACATCGATTACAAGGATGACGATGACAAGATGGCCCCCAAGAAGAAGAGGAAGGTGGGCATTCACGGGGTACCCATGGTGGATCTACGCACGCTCGGCTACAGCCAGCAGCAACAGGAGAAGATCAAACCGAAGGTTCGTTCGACAGTGGCGCAGCACCACGAGGCACTGGTCGGCCATGGGTTTACACACGCGCACATCGTTGCGCTCAGCCAACACCCGGCAGCGTTAGGGACCGTCGCTGTCAAGTATCAGGACATGATCGCAGCGTTGCCAGAGGCGACACACGAAGCGATCGTTGGCGTCGGCAAACAGTGGTCCGGCGCACGCGCCCTGGAGGCCTTGCTCACGGTGGCGGGAGAGTTGAGAGGTCCACCGTTACAGTTGGACACAGGCCAACTTCTCAAGATTGCAAAACGTGGCGGCGTGACCGCAGTGGAGGCAGTGCATGCATGGCGCAATGCACTGACGGGGGCCCCCCTGAACCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCAATATTGGTGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCAATGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCAATATTGGTGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCAATATTGGTGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGCACAGGTGGTGGCCATCGCCAGCAATATTGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGCGCCAGGCCCATGGCCTGACCCCGGAGCAGGTCGTGGCCATCGCCAGCAATGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGCACAGGTGGTGGCCATCGCCAGCAATATTGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTCGTGGCCATCGCCAGCCACGATGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCCACGATGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGACCAGGTGGTGGCCATCGCCAGCAATGGCGGTGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGACCAGGTGGTGGCCATCGCCAGCAATGGCGGTGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCCACGATGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCAATGGCGGCGGCAGGCCGGCGCTGGAGAGCATTGTTGCCCAGTTATCTCGCCCTGATCCGTCGTTGGCCGCGTTAACCAACGACCACCTCGTCGCCTTGGCCTGCCTCGGCGGACGTCCTGCGCTGGATGCAGTGAAAAAGGGATTGCCGCACGCGCCGGCCTTGATCAAAAGAACCGGATCCCAGCTGGTGAAGAGCGAGCTGGAGGAGAAGAAGTCCGAGCTGCGGCACAAGCTGAAGTACGTGCCCCACGAGTACATCGAGCTGATCGAGATCGCCAGGAACAGCACCCAGGACCGCATCCTGGAGATGAAGGTGATGGAGTTCTTCATGAAGGTGTACGGCTACAGGGGAAAGCACCTGGGCGGAAGCAGAAAGCCTGACGGCGCCATCTATACAGTGGGCAGCCCCATCGATTACGGCGTGATCGTGGACACAAAGGCCTACAGCGGCGGCTACAATCTGCCTATCGGCCAGGCCGACGAGATGCAGAGATACGTGGAGGAGAACCAGACCCGGAATAAGCACATCAACCCCAACGAGTGGTGGAAGGTGTACCCTAGCAGCGTGACCGAGTTCAAGTTCCTGTTCGTGAGCGGCCACTTCAAGGGCAACTACAAGGCCCAGCTGACCAGGCTGAACCACATCACCAACTGCAATGGCGCCGTGCTGAGCGTGGAGGAGCTGCTGATCGGCGGCGAGATGATCAAAGCCGGCACCCTGACACTGGAGGAGGTGCGGCGCAAGTTCAACAACGGCGAGATCAACTTCAGATCT
MDYKDHDGDYKDHDIDYKDDDDKMAPKKKRKVGIHGVPMVDLRTLGYSQQQQEKIKPKVRSTVAQHHEALVGHGFTHAHIVALSQHPAALGTVAVKYQDMIAALPEATHEAIVGVGKQWSGARALEALLTVAGELRGPPLQLDTGQLLKIAKRGGVTAVEAVHAWRNALTGAPLNLTPEQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNIGGKQALETVQRLLPVLRQAHGLTPEQVVAIASNGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNGGGRPALESIVAQLSRPDPSLAALTNDHLVALACLGGRPALDAVKKGLPHAPALIKRTNRRIPERTSHRVAGSQLVKSELEEKKSELRHKLKYVPHEYIELIEIARNSTQDRILEMKVMEFFMKVYGYRGKHLGGSRKPDGAIYTVGSPIDYGVIVDTKAYSGGYNLPIGQADEMQRYVEENQTRNKHINPNEWWKVYPSSVTEFKFLFVSGHFKGNYKAQLTRLNHITNCNGAVLSVEELLIGGEMIKAGTLTLEEVRRKFNNGEINFRS
ATGGACTACAAAGACCATGACGGTGATTATAAAGATCATGACATCGATTACAAGGATGACGATGACAAGATGGCCCCCAAGAAGAAGAGGAAGGTGGGCATTCACGGGGTACCCATGGTGGATCTACGCACGCTCGGCTACAGCCAGCAGCAACAGGAGAAGATCAAACCGAAGGTTCGTTCGACAGTGGCGCAGCACCACGAGGCACTGGTCGGCCATGGGTTTACACACGCGCACATCGTTGCGCTCAGCCAACACCCGGCAGCGTTAGGGACCGTCGCTGTCAAGTATCAGGACATGATCGCAGCGTTGCCAGAGGCGACACACGAAGCGATCGTTGGCGTCGGCAAACAGTGGTCCGGCGCACGCGCCCTGGAGGCCTTGCTCACGGTGGCGGGAGAGTTGAGAGGTCCACCGTTACAGTTGGACACAGGCCAACTTCTCAAGATTGCAAAACGTGGCGGCGTGACCGCAGTGGAGGCAGTGCATGCATGGCGCAATGCACTGACGGGGGCCCCCCTGAACCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCAATATTGGTGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCAATGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCAATATTGGTGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCAATATTGGTGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGCACAGGTGGTGGCCATCGCCAGCAATATTGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGCGCCAGGCCCATGGCCTGACCCCGGAGCAGGTCGTGGCCATCGCCAGCAATGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGCACAGGTGGTGGCCATCGCCAGCAATATTGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTCGTGGCCATCGCCAGCCACGATGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCCACGATGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGACCAGGTGGTGGCCATCGCCAGCAATGGCGGTGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGACCAGGTGGTGGCCATCGCCAGCAATGGCGGTGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCCACGATGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCAATGGCGGCGGCAGGCCGGCGCTGGAGAGCATTGTTGCCCAGTTATCTCGCCCTGATCCGTCGTTGGCCGCGTTAACCAACGACCACCTCGTCGCCTTGGCCTGCCTCGGCGGACGTCCTGCGCTGGATGCAGTGAAAAAGGGATTGCCGCACGCGCCGGCCTTGATCAAAAGAACCAATCGCCGTATTCCCGAACGCACATCCCATCGCGTTGCCGGATCCCAGCTGGTGAAGAGCGAGCTGGAGGAGAAGAAGTCCGAGCTGCGGCACAAGCTGAAGTACGTGCCCCACGAGTACATCGAGCTGATCGAGATCGCCAGGAACAGCACCCAGGACCGCATCCTGGAGATGAAGGTGATGGAGTTCTTCATGAAGGTGTACGGCTACAGGGGAAAGCACCTGGGCGGAAGCAGAAAGCCTGACGGCGCCATCTATACAGTGGGCAGCCCCATCGATTACGGCGTGATCGTGGACACAAAGGCCTACAGCGGCGGCTACAATCTGCCTATCGGCCAGGCCGACGAGATGCAGAGATACGTGGAGGAGAACCAGACCCGGAATAAGCACATCAACCCCAACGAGTGGTGGAAGGTGTACCCTAGCAGCGTGACCGAGTTCAAGTTCCTGTTCGTGAGCGGCCACTTCAAGGGCAACTACAAGGCCCAGCTGACCAGGCTGAACCACATCACCAACTGCAATGGCGCCGTGCTGAGCGTGGAGGAGCTGCTGATCGGCGGCGAGATGATCAAAGCCGGCACCCTGACACTGGAGGAGGTGCGGCGCAAGTTCAACAACGGCGAGATCAACTTCAGATCT
MDYKDHDGDYKDHDIDYKDDDDKMAPKKKRKVGIHGVPMVDLRTLGYSQQQQEKIKPKVRSTVAQHHEALVGHGFTHAHIVALSQHPAALGTVAVKYQDMIAALPEATHEAIVGVGKQWSGARALEALLTVAGELRGPPLQLDTGQLLKIAKRGGVTAVEAVHAWRNALTGAPLNLTPEQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNIGGKQALETVQRLLPVLRQAHGLTPEQVVAIASNGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNGGGRPALESIVAQLSRPDPSLAALTNDHLVALACLGGRPALDAVKKGLPHAPALIKRTNRRIPERTSHRVADHAQVVRVLGFFQCHSGSQLVKSELEEKKSELRHKLKYVPHEYIELIEIARNSTQDRILEMKVMEFFMKVYGYRGKHLGGSRKPDGAIYTVGSPIDYGVIVDTKAYSGGYNLPIGQADEMQRYVEENQTRNKHINPNEWWKVYPSSVTEFKFLFVSGHFKGNYKAQLTRLNHITNCNGAVLSVEELLIGGEMIKAGTLTLEEVRRKFNNGEINFRS
ATGGACTACAAAGACCATGACGGTGATTATAAAGATCATGACATCGATTACAAGGATGACGATGACAAGATGGCCCCCAAGAAGAAGAGGAAGGTGGGCATTCACGGGGTACCCATGGTGGATCTACGCACGCTCGGCTACAGCCAGCAGCAACAGGAGAAGATCAAACCGAAGGTTCGTTCGACAGTGGCGCAGCACCACGAGGCACTGGTCGGCCATGGGTTTACACACGCGCACATCGTTGCGCTCAGCCAACACCCGGCAGCGTTAGGGACCGTCGCTGTCAAGTATCAGGACATGATCGCAGCGTTGCCAGAGGCGACACACGAAGCGATCGTTGGCGTCGGCAAACAGTGGTCCGGCGCACGCGCCCTGGAGGCCTTGCTCACGGTGGCGGGAGAGTTGAGAGGTCCACCGTTACAGTTGGACACAGGCCAACTTCTCAAGATTGCAAAACGTGGCGGCGTGACCGCAGTGGAGGCAGTGCATGCATGGCGCAATGCACTGACGGGGGCCCCCCTGAACCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCAATATTGGTGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCAATGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCAATATTGGTGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCAATATTGGTGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGCACAGGTGGTGGCCATCGCCAGCAATATTGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGCGCCAGGCCCATGGCCTGACCCCGGAGCAGGTCGTGGCCATCGCCAGCAATGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGCACAGGTGGTGGCCATCGCCAGCAATATTGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTCGTGGCCATCGCCAGCCACGATGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCCACGATGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGACCAGGTGGTGGCCATCGCCAGCAATGGCGGTGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGACCAGGTGGTGGCCATCGCCAGCAATGGCGGTGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCCACGATGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCAATGGCGGCGGCAGGCCGGCGCTGGAGAGCATTGTTGCCCAGTTATCTCGCCCTGATCCGTCGTTGGCCGCGTTAACCAACGACCACCTCGTCGCCTTGGCCTGCCTCGGCGGACGTCCTGCGCTGGATGCAGTGAAAAAGGGATTGCCGCACGCGCCGGCCTTGATCAAAAGAACCAATCGCCGTATTCCCGAACGCACATCCCATCGCGTTGCCGACCACGCGCAAGTGGTTCGCGTGCTGGGTTTTTTCCAGTGCCACTCCGGATCCCAGCTGGTGAAGAGCGAGCTGGAGGAGAAGAAGTCCGAGCTGCGGCACAAGCTGAAGTACGTGCCCCACGAGTACATCGAGCTGATCGAGATCGCCAGGAACAGCACCCAGGACCGCATCCTGGAGATGAAGGTGATGGAGTTCTTCATGAAGGTGTACGGCTACAGGGGAAAGCACCTGGGCGGAAGCAGAAAGCCTGACGGCGCCATCTATACAGTGGGCAGCCCCATCGATTACGGCGTGATCGTGGACACAAAGGCCTACAGCGGCGGCTACAATCTGCCTATCGGCCAGGCCGACGAGATGCAGAGATACGTGGAGGAGAACCAGACCCGGAATAAGCACATCAACCCCAACGAGTGGTGGAAGGTGTACCCTAGCAGCGTGACCGAGTTCAAGTTCCTGTTCGTGAGCGGCCACTTCAAGGGCAACTACAAGGCCCAGCTGACCAGGCTGAACCACATCACCAACTGCAATGGCGCCGTGCTGAGCGTGGAGGAGCTGCTGATCGGCGGCGAGATGATCAAAGCCGGCACCCTGACACTGGAGGAGGTGCGGCGCAAGTTCAACAACGGCGAGATCAACTTCAGATCT
MDYKDHDGDYKDHDIDYKDDDDKMAPKKKRKVGIHGVPMVDLRTLGYSQQQQEKIKPKVRSTVAQHHEALVGHGFTHAHIVALSQHPAALGTVAVKYQDMIAALPEATHEAIVGVGKQWSGARALEALLTVAGELRGPPLQLDTGQLLKIAKRGGVTAVEAVHAWRNALTGAPLNLTPEQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNIGGKQALETVQRLLPVLRQAHGLTPEQVVAIASNGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNGGGRPALESIVAQLSRPDPSLAALTNDHLVALACLGGRPALDAVKKGLPHAPALIKRTNRRIPERTSHRVADHAQVVRVLGFFQCHSHPAQAFDDAMTQFGMSGSQLVKSELEEKKSELRHKLKYVPHEYIELIEIARNSTQDRILEMKVMEFFMKVYGYRGKHLGGSRKPDGAIYTVGSPIDYGVIVDTKAYSGGYNLPIGQADEMQRYVEENQTRNKHINPNEWWKVYPSSVTEFKFLFVSGHFKGNYKAQLTRLNHITNCNGAVLSVEELLIGGEMIKAGTLTLEEVRRKFNNGEINFRS
ATGGACTACAAAGACCATGACGGTGATTATAAAGATCATGACATCGATTACAAGGATGACGATGACAAGATGGCCCCCAAGAAGAAGAGGAAGGTGGGCATTCACGGGGTACCCATGGTGGATCTACGCACGCTCGGCTACAGCCAGCAGCAACAGGAGAAGATCAAACCGAAGGTTCGTTCGACAGTGGCGCAGCACCACGAGGCACTGGTCGGCCATGGGTTTACACACGCGCACATCGTTGCGCTCAGCCAACACCCGGCAGCGTTAGGGACCGTCGCTGTCAAGTATCAGGACATGATCGCAGCGTTGCCAGAGGCGACACACGAAGCGATCGTTGGCGTCGGCAAACAGTGGTCCGGCGCACGCGCCCTGGAGGCCTTGCTCACGGTGGCGGGAGAGTTGAGAGGTCCACCGTTACAGTTGGACACAGGCCAACTTCTCAAGATTGCAAAACGTGGCGGCGTGACCGCAGTGGAGGCAGTGCATGCATGGCGCAATGCACTGACGGGGGCCCCCCTGAACCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCAATATTGGTGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCAATGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCAATATTGGTGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCAATATTGGTGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGCACAGGTGGTGGCCATCGCCAGCAATATTGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGCGCCAGGCCCATGGCCTGACCCCGGAGCAGGTCGTGGCCATCGCCAGCAATGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGCACAGGTGGTGGCCATCGCCAGCAATATTGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTCGTGGCCATCGCCAGCCACGATGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCCACGATGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGACCAGGTGGTGGCCATCGCCAGCAATGGCGGTGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGACCAGGTGGTGGCCATCGCCAGCAATGGCGGTGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCCACGATGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCAATGGCGGCGGCAGGCCGGCGCTGGAGAGCATTGTTGCCCAGTTATCTCGCCCTGATCCGTCGTTGGCCGCGTTAACCAACGACCACCTCGTCGCCTTGGCCTGCCTCGGCGGACGTCCTGCGCTGGATGCAGTGAAAAAGGGATTGCCGCACGCGCCGGCCTTGATCAAAAGAACCAATCGCCGTATTCCCGAACGCACATCCCATCGCGTTGCCGACCACGCGCAAGTGGTTCGCGTGCTGGGTTTTTTCCAGTGCCACTCCCACCCAGCGCAAGCATTTGATGACGCCATGACGCAGTTCGGGATGAGCGGATCCCAGCTGGTGAAGAGCGAGCTGGAGGAGAAGAAGTCCGAGCTGCGGCACAAGCTGAAGTACGTGCCCCACGAGTACATCGAGCTGATCGAGATCGCCAGGAACAGCACCCAGGACCGCATCCTGGAGATGAAGGTGATGGAGTTCTTCATGAAGGTGTACGGCTACAGGGGAAAGCACCTGGGCGGAAGCAGAAAGCCTGACGGCGCCATCTATACAGTGGGCAGCCCCATCGATTACGGCGTGATCGTGGACACAAAGGCCTACAGCGGCGGCTACAATCTGCCTATCGGCCAGGCCGACGAGATGCAGAGATACGTGGAGGAGAACCAGACCCGGAATAAGCACATCAACCCCAACGAGTGGTGGAAGGTGTACCCTAGCAGCGTGACCGAGTTCAAGTTCCTGTTCGTGAGCGGCCACTTCAAGGGCAACTACAAGGCCCAGCTGACCAGGCTGAACCACATCACCAACTGCAATGGCGCCGTGCTGAGCGTGGAGGAGCTGCTGATCGGCGGCGAGATGATCAAAGCCGGCACCCTGACACTGGAGGAGGTGCGGCGCAAGTTCAACAACGGCGAGATCAACTTCAGATCT
コード配列に下線を引いた、完全なTALEN構築物配列(配列番号238):
GACTCTTCGCGATGTACGGGCCAGATATACGCGTTGACATTGATTATTGACTAGTTATTAATAGTAATCAATTACGGGGTCATTAGTTCATAGCCCATATATGGAGTTCCGCGTTACATAACTTACGGTAAATGGCCCGCCTGGCTGACCGCCCAACGACCCCCGCCCATTGACGTCAATAATGACGTATGTTCCCATAGTAACGCCAATAGGGACTTTCCATTGACGTCAATGGGTGGAGTATTTACGGTAAACTGCCCACTTGGCAGTACATCAAGTGTATCATATGCCAAGTACGCCCCCTATTGACGTCAATGACGGTAAATGGCCCGCCTGGCATTATGCCCAGTACATGACCTTATGGGACTTTCCTACTTGGCAGTACATCTACGTATTAGTCATCGCTATTACCATGGTGATGCGGTTTTGGCAGTACATCAATGGGCGTGGATAGCGGTTTGACTCACGGGGATTTCCAAGTCTCCACCCCATTGACGTCAATGGGAGTTTGTTTTGGCACCAAAATCAACGGGACTTTCCAAAATGTCGTAACAACTCCGCCCCATTGACGCAAATGGGCGGTAGGCGTGTACGGTGGGAGGTCTATATAAGCAGAGCTCTCTGGCTAACTAGAGAACCCACTGCTTACTGGCTTATCGAAATTAATACGACTCACTATAGGGAGAGCCAAGCTGACTAGCGTTTAAACTTAAGCTGATCCACTAGTCCAGTGTGGTGGAATTCGCCATGGACTACAAAGACCATGACGGTGATTATAAAGATCATGACATCGATTACAAGGATGACGATGACAAGATGGCCCCCAAGAAGAAGAGGAAGGTGGGCATTCACCGCGGGGTACCTATGGTGGACTTGAGGACACTCGGTTATTCGCAACAGCAACAGGAGAAAATCAAGCCTAAGGTCAGGAGCACCGTCGCGCAACACCACGAGGCGCTTGTGGGGCATGGCTTCACTCATGCGCATATTGTCGCGCTTTCACAGCACCCTGCGGCGCTTGGGACGGTGGCTGTCAAATACCAAGATATGATTGCGGCCCTGCCCGAAGCCACGCACGAGGCAATTGTAGGGGTCGGTAAACAGTGGTCGGGAGCGCGAGCACTTGAGGCGCTGCTGACTGTGGCGGGTGAGCTTAGGGGGCCTCCGCTCCAGCTCGACACCGGGCAGCTGCTGAAGATCGCGAAGAGAGGGGGAGTAACAGCGGTAGAGGCAGTGCACGCCTGGCGCAATGCGCTCACCGGGGCCCCCTTGAACCTGACCCCAGACCAGGTAGTCGCAATCGCCAACAATAACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCACATGACGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCAATGGGGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCAACAACAACGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGACTCACCCCAGACCAGGTAGTCGCAATCGCCAACAATAACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCAAACGGAGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCCATGATGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCAATATTGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGGCTGACCCCAGACCAGGTAGTCGCAATCGCGTCGAATGGCGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCACATGACGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCCATGATGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCAACGGTGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGTTTGACCCCAGACCAGGTAGTCGCAATCGCGTCGCATGACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACGCCTGCACAAGTGGTCGCCATCGCCTCCAATATTGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGCCTGACACCCGAACAGGTGGTCGCCATTGCTAGCAACGGGGGAGGACGGCCAGCCTTGGAGTCCATCGTAGCCCAATTGTCCAGGCCCGATCCCGCGTTGGCTGCGTTAACGAATGACCATCTGGTGGCGTTGGCATGTCTTGGTGGATCCCAGCTGGTGAAGAGCGAGCTGGAGGAGAAGAAGTCCGAGCTGCGGCACAAGCTGAAGTACGTGCCCCACGAGTACATCGAGCTGATCGAGATCGCCAGGAACAGCACCCAGGACCGCATCCTGGAGATGAAGGTGATGGAGTTCTTCATGAAGGTGTACGGCTACAGGGGAAAGCACCTGGGCGGAAGCAGAAAGCCTGACGGCGCCATCTATACAGTGGGCAGCCCCATCGATTACGGCGTGATCGTGGACACAAAGGCCTACAGCGGCGGCTACAATCTGCCTATCGGCCAGGCCGACGAGATGCAGAGATACGTGGAGGAGAACCAGACCCGGAATAAGCACATCAACCCCAACGAGTGGTGGAAGGTGTACCCTAGCAGCGTGACCGAGTTCAAGTTCCTGTTCGTGAGCGGCCACTTCAAGGGCAACTACAAGGCCCAGCTGACCAGGCTGAACCACATCACCAACTGCAATGGCGCCGTGCTGAGCGTGGAGGAGCTGCTGATCGGCGGCGAGATGATCAAAGCCGGCACCCTGACACTGGAGGAGGTGCGGCGCAAGTTCAACAACGGCGAGATCAACTTCAGATCTTGATAACTCGAGTCTAGAGGGCCCGTTTAAACCCGCTGATCAGCCTCGACTGTGCCTTCTAGTTGCCAGCCATCTGTTGTTTGCCCCTCCCCCGTGCCTTCCTTGACCCTGGAAGGTGCCACTCCCACTGTCCTTTCCTAATAAAATGAGGAAATTGCATCGCATTGTCTGAGTAGGTGTCATTCTATTCTGGGGGGTGGGGTGGGGCAGGACAGCAAGGGGGAGGATTGGGAAGACAATAGCAGGCATGCTGGGGATGCGGTGGGCTCTATGGCTTCTACTGGGCGGTTTTATGGACAGCAAGCGAACCGGAATTGCCAGCTGGGGCGCCCTCTGGTAAGGTTGGGAAGCCCTGCAAAGTAAACTGGATGGCTTTCTCGCCGCCAAGGATCTGATGGCGCAGGGGATCAAGCTCTGATCAAGAGACAGGATGAGGATCGTTTCGCATGATTGAACAAGATGGATTGCACGCAGGTTCTCCGGCCGCTTGGGTGGAGAGGCTATTCGGCTATGACTGGGCACAACAGACAATCGGCTGCTCTGATGCCGCCGTGTTCCGGCTGTCAGCGCAGGGGCGCCCGGTTCTTTTTGTCAAGACCGACCTGTCCGGTGCCCTGAATGAACTGCAAGACGAGGCAGCGCGGCTATCGTGGCTGGCCACGACGGGCGTTCCTTGCGCAGCTGTGCTCGACGTTGTCACTGAAGCGGGAAGGGACTGGCTGCTATTGGGCGAAGTGCCGGGGCAGGATCTCCTGTCATCTCACCTTGCTCCTGCCGAGAAAGTATCCATCATGGCTGATGCAATGCGGCGGCTGCATACGCTTGATCCGGCTACCTGCCCATTCGACCACCAAGCGAAACATCGCATCGAGCGAGCACGTACTCGGATGGAAGCCGGTCTTGTCGATCAGGATGATCTGGACGAAGAGCATCAGGGGCTCGCGCCAGCCGAACTGTTCGCCAGGCTCAAGGCGAGCATGCCCGACGGCGAGGATCTCGTCGTGACCCATGGCGATGCCTGCTTGCCGAATATCATGGTGGAAAATGGCCGCTTTTCTGGATTCATCGACTGTGGCCGGCTGGGTGTGGCGGACCGCTATCAGGACATAGCGTTGGCTACCCGTGATATTGCTGAAGAGCTTGGCGGCGAATGGGCTGACCGCTTCCTCGTGCTTTACGGTATCGCCGCTCCCGATTCGCAGCGCATCGCCTTCTATCGCCTTCTTGACGAGTTCTTCTGAATTATTAACGCTTACAATTTCCTGATGCGGTATTTTCTCCTTACGCATCTGTGCGGTATTTCACACCGCATACAGGTGGCACTTTTCGGGGAAATGTGCGCGGAACCCCTATTTGTTTATTTTTCTAAATACATTCAAATATGTATCCGCTCATGAGACAATAACCCTGATAAATGCTTCAATAATAGCACGTGCTAAAACTTCATTTTTAATTTAAAAGGATCTAGGTGAAGATCCTTTTTGATAATCTCATGACCAAAATCCCTTAACGTGAGTTTTCGTTCCACTGAGCGTCAGACCCCGTAGAAAAGATCAAAGGATCTTCTTGAGATCCTTTTTTTCTGCGCGTAATCTGCTGCTTGCAAACAAAAAAACCACCGCTACCAGCGGTGGTTTGTTTGCCGGATCAAGAGCTACCAACTCTTTTTCCGAAGGTAACTGGCTTCAGCAGAGCGCAGATACCAAATACTGTCCTTCTAGTGTAGCCGTAGTTAGGCCACCACTTCAAGAACTCTGTAGCACCGCCTACATACCTCGCTCTGCTAATCCTGTTACCAGTGGCTGCTGCCAGTGGCGATAAGTCGTGTCTTACCGGGTTGGACTCAAGACGATAGTTACCGGATAAGGCGCAGCGGTCGGGCTGAACGGGGGGTTCGTGCACACAGCCCAGCTTGGAGCGAACGACCTACACCGAACTGAGATACCTACAGCGTGAGCTATGAGAAAGCGCCACGCTTCCCGAAGGGAGAAAGGCGGACAGGTATCCGGTAAGCGGCAGGGTCGGAACAGGAGAGCGCACGAGGGAGCTTCCAGGGGGAAACGCCTGGTATCTTTATAGTCCTGTCGGGTTTCGCCACCTCTGACTTGAGCGTCGATTTTTGTGATGCTCGTCAGGGGGGCGGAGCCTATGGAAAAACGCCAGCAACGCGGCCTTTTTACGGTTCCTGGGCTTTTGCTGGCCTTTTGCTCACATGTTCTT
それぞれの発現構築物の配列を再生成するために、上述の構築物の下線を引いた領域を以下に示されるそれぞれのCDSに置換する。
>CCR5 L161(+28)(配列番号239)
MDYKDHDGDYKDHDIDYKDDDDKMAPKKKRKVGIHRGVPMVDLRTLGYSQQQQEKIKPKVRSTVAQHHEALVGHGFTHAHIVALSQHPAALGTVAVKYQDMIAALPEATHEAIVGVGKQWSGARALEALLTVAGELRGPPLQLDTGQLLKIAKRGGVTAVEAVHAWRNALTGAPLNLTPDQVVAIANNNGGKQALETVQRLLPVLCQDHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPAQVVAIANNNGGKQALETVQRLLPVLCQDHGLTPDQVVAIANNNGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQDHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNGGGKQALETVQRLLPVLCQDHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPAQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGRPALESIVAQLSRPDPALAALTNDHLVALACLGGSQLVKSELEEKKSELRHKLKYVPHEYIELIEIARNSTQDRILEMKVMEFFMKVYGYRGKHLGGSRKPDGAIYTVGSPIDYGVIVDTKAYSGGYNLPIGQADEMQRYVEENQTRNKHINPNEWWKVYPSSVTEFKFLFVSGHFKGNYKAQLTRLNHITNCNGAVLSVEELLIGGEMIKAGTLTLEEVRRKFNNGEINFRS
ATGGACTACAAAGACCATGACGGTGATTATAAAGATCATGACATCGATTACAAGGATGACGATGACAAGATGGCCCCCAAGAAGAAGAGGAAGGTGGGCATTCACCGCGGGGTACCTATGGTGGACTTGAGGACACTCGGTTATTCGCAACAGCAACAGGAGAAAATCAAGCCTAAGGTCAGGAGCACCGTCGCGCAACACCACGAGGCGCTTGTGGGGCATGGCTTCACTCATGCGCATATTGTCGCGCTTTCACAGCACCCTGCGGCGCTTGGGACGGTGGCTGTCAAATACCAAGATATGATTGCGGCCCTGCCCGAAGCCACGCACGAGGCAATTGTAGGGGTCGGTAAACAGTGGTCGGGAGCGCGAGCACTTGAGGCGCTGCTGACTGTGGCGGGTGAGCTTAGGGGGCCTCCGCTCCAGCTCGACACCGGGCAGCTGCTGAAGATCGCGAAGAGAGGGGGAGTAACAGCGGTAGAGGCAGTGCACGCCTGGCGCAATGCGCTCACCGGGGCCCCCTTGAACCTGACCCCAGACCAGGTAGTCGCAATCGCCAACAATAACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCACATGACGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCAATGGGGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCAACAACAACGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGACTCACCCCAGACCAGGTAGTCGCAATCGCCAACAATAACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCAAACGGAGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCCATGATGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCAATATTGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGGCTGACCCCAGACCAGGTAGTCGCAATCGCGTCGAATGGCGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCACATGACGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCCATGATGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCAACGGTGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGTTTGACCCCAGACCAGGTAGTCGCAATCGCGTCGCATGACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACGCCTGCACAAGTGGTCGCCATCGCCTCCAATATTGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGCCTGACACCCGAACAGGTGGTCGCCATTGCTAGCAACGGGGGAGGACGGCCAGCCTTGGAGTCCATCGTAGCCCAATTGTCCAGGCCCGATCCCGCGTTGGCTGCGTTAACGAATGACCATCTGGTGGCGTTGGCATGTCTTGGTGGATCCCAGCTGGTGAAGAGCGAGCTGGAGGAGAAGAAGTCCGAGCTGCGGCACAAGCTGAAGTACGTGCCCCACGAGTACATCGAGCTGATCGAGATCGCCAGGAACAGCACCCAGGACCGCATCCTGGAGATGAAGGTGATGGAGTTCTTCATGAAGGTGTACGGCTACAGGGGAAAGCACCTGGGCGGAAGCAGAAAGCCTGACGGCGCCATCTATACAGTGGGCAGCCCCATCGATTACGGCGTGATCGTGGACACAAAGGCCTACAGCGGCGGCTACAATCTGCCTATCGGCCAGGCCGACGAGATGCAGAGATACGTGGAGGAGAACCAGACCCGGAATAAGCACATCAACCCCAACGAGTGGTGGAAGGTGTACCCTAGCAGCGTGACCGAGTTCAAGTTCCTGTTCGTGAGCGGCCACTTCAAGGGCAACTACAAGGCCCAGCTGACCAGGCTGAACCACATCACCAACTGCAATGGCGCCGTGCTGAGCGTGGAGGAGCTGCTGATCGGCGGCGAGATGATCAAAGCCGGCACCCTGACACTGGAGGAGGTGCGGCGCAAGTTCAACAACGGCGAGATCAACTTCAGATCT
MDYKDHDGDYKDHDIDYKDDDDKMAPKKKRKVGIHRGVPMVDLRTLGYSQQQQEKIKPKVRSTVAQHHEALVGHGFTHAHIVALSQHPAALGTVAVKYQDMIAALPEATHEAIVGVGKQWSGARALEALLTVAGELRGPPLQLDTGQLLKIAKRGGVTAVEAVHAWRNALTGAPLNLTPDQVVAIANNNGGKQALETVQRLLPVLCQDHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPAQVVAIANNNGGKQALETVQRLLPVLCQDHGLTPDQVVAIANNNGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQDHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNGGGKQALETVQRLLPVLCQDHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPAQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGRPALESIVAQLSRPDPALAALTNDHLVALACLGGRPALDAVKKGLPHAPALIKRTNRRIPERTSHRVAGSQLVKSELEEKKSELRHKLKYVPHEYIELIEIARNSTQDRILEMKVMEFFMKVYGYRGKHLGGSRKPDGAIYTVGSPIDYGVIVDTKAYSGGYNLPIGQADEMQRYVEENQTRNKHINPNEWWKVYPSSVTEFKFLFVSGHFKGNYKAQLTRLNHITNCNGAVLSVEELLIGGEMIKAGTLTLEEVRRKFNNGEINFRS
ATGGACTACAAAGACCATGACGGTGATTATAAAGATCATGACATCGATTACAAGGATGACGATGACAAGATGGCCCCCAAGAAGAAGAGGAAGGTGGGCATTCACCGCGGGGTACCTATGGTGGACTTGAGGACACTCGGTTATTCGCAACAGCAACAGGAGAAAATCAAGCCTAAGGTCAGGAGCACCGTCGCGCAACACCACGAGGCGCTTGTGGGGCATGGCTTCACTCATGCGCATATTGTCGCGCTTTCACAGCACCCTGCGGCGCTTGGGACGGTGGCTGTCAAATACCAAGATATGATTGCGGCCCTGCCCGAAGCCACGCACGAGGCAATTGTAGGGGTCGGTAAACAGTGGTCGGGAGCGCGAGCACTTGAGGCGCTGCTGACTGTGGCGGGTGAGCTTAGGGGGCCTCCGCTCCAGCTCGACACCGGGCAGCTGCTGAAGATCGCGAAGAGAGGGGGAGTAACAGCGGTAGAGGCAGTGCACGCCTGGCGCAATGCGCTCACCGGGGCCCCCTTGAACCTGACCCCAGACCAGGTAGTCGCAATCGCCAACAATAACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCACATGACGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCAATGGGGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCAACAACAACGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGACTCACCCCAGACCAGGTAGTCGCAATCGCCAACAATAACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCAAACGGAGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCCATGATGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCAATATTGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGGCTGACCCCAGACCAGGTAGTCGCAATCGCGTCGAATGGCGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCACATGACGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCCATGATGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCAACGGTGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGTTTGACCCCAGACCAGGTAGTCGCAATCGCGTCGCATGACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACGCCTGCACAAGTGGTCGCCATCGCCTCCAATATTGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGCCTGACACCCGAACAGGTGGTCGCCATTGCTAGCAACGGGGGAGGACGGCCAGCCTTGGAGTCCATCGTAGCCCAATTGTCCAGGCCCGATCCCGCGTTGGCTGCGTTAACGAATGACCATCTGGTGGCGTTGGCATGTCTTGGTGGACGACCCGCGCTCGATGCAGTCAAAAAGGGTCTGCCTCATGCTCCCGCATTGATCAAAAGAACCAACCGGCGGATTCCCGAGAGAACTTCCCATCGAGTCGCGGGATCCCAGCTGGTGAAGAGCGAGCTGGAGGAGAAGAAGTCCGAGCTGCGGCACAAGCTGAAGTACGTGCCCCACGAGTACATCGAGCTGATCGAGATCGCCAGGAACAGCACCCAGGACCGCATCCTGGAGATGAAGGTGATGGAGTTCTTCATGAAGGTGTACGGCTACAGGGGAAAGCACCTGGGCGGAAGCAGAAAGCCTGACGGCGCCATCTATACAGTGGGCAGCCCCATCGATTACGGCGTGATCGTGGACACAAAGGCCTACAGCGGCGGCTACAATCTGCCTATCGGCCAGGCCGACGAGATGCAGAGATACGTGGAGGAGAACCAGACCCGGAATAAGCACATCAACCCCAACGAGTGGTGGAAGGTGTACCCTAGCAGCGTGACCGAGTTCAAGTTCCTGTTCGTGAGCGGCCACTTCAAGGGCAACTACAAGGCCCAGCTGACCAGGCTGAACCACATCACCAACTGCAATGGCGCCGTGCTGAGCGTGGAGGAGCTGCTGATCGGCGGCGAGATGATCAAAGCCGGCACCCTGACACTGGAGGAGGTGCGGCGCAAGTTCAACAACGGCGAGATCAACTTCAGATCT
MDYKDHDGDYKDHDIDYKDDDDKMAPKKKRKVGIHRGVPMVDLRTLGYSQQQQEKIKPKVRSTVAQHHEALVGHGFTHAHIVALSQHPAALGTVAVKYQDMIAALPEATHEAIVGVGKQWSGARALEALLTVAGELRGPPLQLDTGQLLKIAKRGGVTAVEAVHAWRNALTGAPLNLTPDQVVAIANNNGGKQALETVQRLLPVLCQDHGLTPDQVVAIANNNGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPAQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPDQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPAQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQDHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNGGGKQALETVQRLLPVLCQDHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPAQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGRPALESIVAQLSRPDPALAALTNDHLVALACLGGSQLVKSELEEKKSELRHKLKYVPHEYIELIEIARNSTQDRILEMKVMEFFMKVYGYRGKHLGGSRKPDGAIYTVGSPIDYGVIVDTKAYSGGYNLPIGQADEMQRYVEENQTRNKHINPNEWWKVYPSSVTEFKFLFVSGHFKGNYKAQLTRLNHITNCNGAVLSVEELLIGGEMIKAGTLTLEEVRRKFNNGEINFRS
ATGGACTACAAAGACCATGACGGTGATTATAAAGATCATGACATCGATTACAAGGATGACGATGACAAGATGGCCCCCAAGAAGAAGAGGAAGGTGGGCATTCACCGCGGGGTACCTATGGTGGACTTGAGGACACTCGGTTATTCGCAACAGCAACAGGAGAAAATCAAGCCTAAGGTCAGGAGCACCGTCGCGCAACACCACGAGGCGCTTGTGGGGCATGGCTTCACTCATGCGCATATTGTCGCGCTTTCACAGCACCCTGCGGCGCTTGGGACGGTGGCTGTCAAATACCAAGATATGATTGCGGCCCTGCCCGAAGCCACGCACGAGGCAATTGTAGGGGTCGGTAAACAGTGGTCGGGAGCGCGAGCACTTGAGGCGCTGCTGACTGTGGCGGGTGAGCTTAGGGGGCCTCCGCTCCAGCTCGACACCGGGCAGCTGCTGAAGATCGCGAAGAGAGGGGGAGTAACAGCGGTAGAGGCAGTGCACGCCTGGCGCAATGCGCTCACCGGGGCCCCCTTGAACCTGACCCCAGACCAGGTAGTCGCAATCGCCAACAATAACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGATCAAGTCGTGGCCATTGCAAATAATAACGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCAATGGGGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCCACGACGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGACCATGGACTCACCCCAGACCAGGTAGTCGCAATCGCGTCGAACATTGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCAAACGGAGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCCATGATGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCCACGACGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGGCTGACCCCAGACCAGGTAGTCGCAATCGCGTCGAATGGCGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCACATGACGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCAACATCGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCAACGGTGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGTTTGACCCCAGACCAGGTAGTCGCAATCGCGTCGCATGACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACGCCTGCACAAGTGGTCGCCATCGCCTCCCACGACGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGCCTGACACCCGAACAGGTGGTCGCCATTGCTAGCAACGGGGGAGGACGGCCAGCCTTGGAGTCCATCGTAGCCCAATTGTCCAGGCCCGATCCCGCGTTGGCTGCGTTAACGAATGACCATCTGGTGGCGTTGGCATGTCTTGGTGGATCCCAGCTGGTGAAGAGCGAGCTGGAGGAGAAGAAGTCCGAGCTGCGGCACAAGCTGAAGTACGTGCCCCACGAGTACATCGAGCTGATCGAGATCGCCAGGAACAGCACCCAGGACCGCATCCTGGAGATGAAGGTGATGGAGTTCTTCATGAAGGTGTACGGCTACAGGGGAAAGCACCTGGGCGGAAGCAGAAAGCCTGACGGCGCCATCTATACAGTGGGCAGCCCCATCGATTACGGCGTGATCGTGGACACAAAGGCCTACAGCGGCGGCTACAATCTGCCTATCGGCCAGGCCGACGAGATGCAGAGATACGTGGAGGAGAACCAGACCCGGAATAAGCACATCAACCCCAACGAGTGGTGGAAGGTGTACCCTAGCAGCGTGACCGAGTTCAAGTTCCTGTTCGTGAGCGGCCACTTCAAGGGCAACTACAAGGCCCAGCTGACCAGGCTGAACCACATCACCAACTGCAATGGCGCCGTGCTGAGCGTGGAGGAGCTGCTGATCGGCGGCGAGATGATCAAAGCCGGCACCCTGACACTGGAGGAGGTGCGGCGCAAGTTCAACAACGGCGAGATCAACTTCAGATCT
MDYKDHDGDYKDHDIDYKDDDDKMAPKKKRKVGIHRGVPMVDLRTLGYSQQQQEKIKPKVRSTVAQHHEALVGHGFTHAHIVALSQHPAALGTVAVKYQDMIAALPEATHEAIVGVGKQWSGARALEALLTVAGELRGPPLQLDTGQLLKIAKRGGVTAVEAVHAWRNALTGAPLNLTPDQVVAIANNNGGKQALETVQRLLPVLCQDHGLTPDQVVAIANNNGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPAQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPDQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPAQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQDHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNGGGKQALETVQRLLPVLCQDHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPAQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGRPALESIVAQLSRPDPALAALTNDHLVALACLGGRPALDAVKKGLPHAPALIKRTNRRIPERTSHRVAGSQLVKSELEEKKSELRHKLKYVPHEYIELIEIARNSTQDRILEMKVMEFFMKVYGYRGKHLGGSRKPDGAIYTVGSPIDYGVIVDTKAYSGGYNLPIGQADEMQRYVEENQTRNKHINPNEWWKVYPSSVTEFKFLFVSGHFKGNYKAQLTRLNHITNCNGAVLSVEELLIGGEMIKAGTLTLEEVRRKFNNGEINFRS
ATGGACTACAAAGACCATGACGGTGATTATAAAGATCATGACATCGATTACAAGGATGACGATGACAAGATGGCCCCCAAGAAGAAGAGGAAGGTGGGCATTCACCGCGGGGTACCTATGGTGGACTTGAGGACACTCGGTTATTCGCAACAGCAACAGGAGAAAATCAAGCCTAAGGTCAGGAGCACCGTCGCGCAACACCACGAGGCGCTTGTGGGGCATGGCTTCACTCATGCGCATATTGTCGCGCTTTCACAGCACCCTGCGGCGCTTGGGACGGTGGCTGTCAAATACCAAGATATGATTGCGGCCCTGCCCGAAGCCACGCACGAGGCAATTGTAGGGGTCGGTAAACAGTGGTCGGGAGCGCGAGCACTTGAGGCGCTGCTGACTGTGGCGGGTGAGCTTAGGGGGCCTCCGCTCCAGCTCGACACCGGGCAGCTGCTGAAGATCGCGAAGAGAGGGGGAGTAACAGCGGTAGAGGCAGTGCACGCCTGGCGCAATGCGCTCACCGGGGCCCCCTTGAACCTGACCCCAGACCAGGTAGTCGCAATCGCCAACAATAACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGATCAAGTCGTGGCCATTGCAAATAATAACGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCAATGGGGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCCACGACGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGACTCACCCCAGACCAGGTAGTCGCAATCGCGTCGAACATTGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCAAACGGAGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCCATGATGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCCACGACGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGGCTGACCCCAGACCAGGTAGTCGCAATCGCGTCGAATGGCGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCACATGACGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCAACATCGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCAACGGTGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGTTTGACCCCAGACCAGGTAGTCGCAATCGCGTCGCATGACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACGCCTGCACAAGTGGTCGCCATCGCCTCCCACGACGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGCCTGACACCCGAACAGGTGGTCGCCATTGCTAGCAACGGGGGAGGACGGCCAGCCTTGGAGTCCATCGTAGCCCAATTGTCCAGGCCCGATCCCGCGTTGGCTGCGTTAACGAATGACCATCTGGTGGCGTTGGCATGTCTTGGTGGACGACCCGCGCTCGATGCAGTCAAAAAGGGTCTGCCTCATGCTCCCGCATTGATCAAAAGAACCAACCGGCGGATTCCCGAGAGAACTTCCCATCGAGTCGCGGGATCCCAGCTGGTGAAGAGCGAGCTGGAGGAGAAGAAGTCCGAGCTGCGGCACAAGCTGAAGTACGTGCCCCACGAGTACATCGAGCTGATCGAGATCGCCAGGAACAGCACCCAGGACCGCATCCTGGAGATGAAGGTGATGGAGTTCTTCATGAAGGTGTACGGCTACAGGGGAAAGCACCTGGGCGGAAGCAGAAAGCCTGACGGCGCCATCTATACAGTGGGCAGCCCCATCGATTACGGCGTGATCGTGGACACAAAGGCCTACAGCGGCGGCTACAATCTGCCTATCGGCCAGGCCGACGAGATGCAGAGATACGTGGAGGAGAACCAGACCCGGAATAAGCACATCAACCCCAACGAGTGGTGGAAGGTGTACCCTAGCAGCGTGACCGAGTTCAAGTTCCTGTTCGTGAGCGGCCACTTCAAGGGCAACTACAAGGCCCAGCTGACCAGGCTGAACCACATCACCAACTGCAATGGCGCCGTGCTGAGCGTGGAGGAGCTGCTGATCGGCGGCGAGATGATCAAAGCCGGCACCCTGACACTGGAGGAGGTGCGGCGCAAGTTCAACAACGGCGAGATCAACTTCAGATCT
MDYKDHDGDYKDHDIDYKDDDDKMAPKKKRKVGIHRGVPMVDLRTLGYSQQQQEKIKPKVRSTVAQHHEALVGHGFTHAHIVALSQHPAALGTVAVKYQDMIAALPEATHEAIVGVGKQWSGARALEALLTVAGELRGPPLQLDTGQLLKIAKRGGVTAVEAVHAWRNALTGAPLNLTPDQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPAQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQDHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNGGGKQALETVQRLLPVLCQDHGLTPDQVVAIANNNGGKQALETVQRLLPVLCQDHGLTPAQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGRPALESIVAQLSRPDPALAALTNDHLVALACLGGSQLVKSELEEKKSELRHKLKYVPHEYIELIEIARNSTQDRILEMKVMEFFMKVYGYRGKHLGGSRKPDGAIYTVGSPIDYGVIVDTKAYSGGYNLPIGQADEMQRYVEENQTRNKHINPNEWWKVYPSSVTEFKFLFVSGHFKGNYKAQLTRLNHITNCNGAVLSVEELLIGGEMIKAGTLTLEEVRRKFNNGEINFRS
ATGGACTACAAAGACCATGACGGTGATTATAAAGATCATGACATCGATTACAAGGATGACGATGACAAGATGGCCCCCAAGAAGAAGAGGAAGGTGGGCATTCACCGCGGGGTACCTATGGTGGACTTGAGGACACTCGGTTATTCGCAACAGCAACAGGAGAAAATCAAGCCTAAGGTCAGGAGCACCGTCGCGCAACACCACGAGGCGCTTGTGGGGCATGGCTTCACTCATGCGCATATTGTCGCGCTTTCACAGCACCCTGCGGCGCTTGGGACGGTGGCTGTCAAATACCAAGATATGATTGCGGCCCTGCCCGAAGCCACGCACGAGGCAATTGTAGGGGTCGGTAAACAGTGGTCGGGAGCGCGAGCACTTGAGGCGCTGCTGACTGTGGCGGGTGAGCTTAGGGGGCCTCCGCTCCAGCTCGACACCGGGCAGCTGCTGAAGATCGCGAAGAGAGGGGGAGTAACAGCGGTAGAGGCAGTGCACGCCTGGCGCAATGCGCTCACCGGGGCCCCCTTGAACCTGACCCCAGACCAGGTAGTCGCAATCGCGTCGCATGACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCAAATATCGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCAATGGGGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCCACGACGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGACTCACCCCAGACCAGGTAGTCGCAATCGCGTCGCATGACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCAAACGGAGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCCATGATGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCAATATTGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGGCTGACCCCAGACCAGGTAGTCGCAATCGCGTCGAATGGCGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCACATGACGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCCATGATGGAGGGAAACAGGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCAACGGTGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGTTTGACCCCAGACCAGGTAGTCGCAATCGCCAACAATAACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACGCCTGCACAAGTGGTCGCCATCGCCTCCAATATTGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGCCTGACACCCGAACAGGTGGTCGCCATTGCTAGCAACGGGGGAGGACGGCCAGCCTTGGAGTCCATCGTAGCCCAATTGTCCAGGCCCGATCCCGCGTTGGCTGCGTTAACGAATGACCATCTGGTGGCGTTGGCATGTCTTGGTGGATCCCAGCTGGTGAAGAGCGAGCTGGAGGAGAAGAAGTCCGAGCTGCGGCACAAGCTGAAGTACGTGCCCCACGAGTACATCGAGCTGATCGAGATCGCCAGGAACAGCACCCAGGACCGCATCCTGGAGATGAAGGTGATGGAGTTCTTCATGAAGGTGTACGGCTACAGGGGAAAGCACCTGGGCGGAAGCAGAAAGCCTGACGGCGCCATCTATACAGTGGGCAGCCCCATCGATTACGGCGTGATCGTGGACACAAAGGCCTACAGCGGCGGCTACAATCTGCCTATCGGCCAGGCCGACGAGATGCAGAGATACGTGGAGGAGAACCAGACCCGGAATAAGCACATCAACCCCAACGAGTGGTGGAAGGTGTACCCTAGCAGCGTGACCGAGTTCAAGTTCCTGTTCGTGAGCGGCCACTTCAAGGGCAACTACAAGGCCCAGCTGACCAGGCTGAACCACATCACCAACTGCAATGGCGCCGTGCTGAGCGTGGAGGAGCTGCTGATCGGCGGCGAGATGATCAAAGCCGGCACCCTGACACTGGAGGAGGTGCGGCGCAAGTTCAACAACGGCGAGATCAACTTCAGATCT
MDYKDHDGDYKDHDIDYKDDDDKMAPKKKRKVGIHRGVPMVDLRTLGYSQQQQEKIKPKVRSTVAQHHEALVGHGFTHAHIVALSQHPAALGTVAVKYQDMIAALPEATHEAIVGVGKQWSGARALEALLTVAGELRGPPLQLDTGQLLKIAKRGGVTAVEAVHAWRNALTGAPLNLTPDQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPAQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQDHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNGGGKQALETVQRLLPVLCQDHGLTPDQVVAIANNNGGKQALETVQRLLPVLCQDHGLTPAQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGRPALESIVAQLSRPDPALAALTNDHLVALACLGGRPALDAVKKGLPHAPALIKRTNRRIPERTSHRVAGSQLVKSELEEKKSELRHKLKYVPHEYIELIEIARNSTQDRILEMKVMEFFMKVYGYRGKHLGGSRKPDGAIYTVGSPIDYGVIVDTKAYSGGYNLPIGQADEMQRYVEENQTRNKHINPNEWWKVYPSSVTEFKFLFVSGHFKGNYKAQLTRLNHITNCNGAVLSVEELLIGGEMIKAGTLTLEEVRRKFNNGEINFRS
ATGGACTACAAAGACCATGACGGTGATTATAAAGATCATGACATCGATTACAAGGATGACGATGACAAGATGGCCCCCAAGAAGAAGAGGAAGGTGGGCATTCACCGCGGGGTACCTATGGTGGACTTGAGGACACTCGGTTATTCGCAACAGCAACAGGAGAAAATCAAGCCTAAGGTCAGGAGCACCGTCGCGCAACACCACGAGGCGCTTGTGGGGCATGGCTTCACTCATGCGCATATTGTCGCGCTTTCACAGCACCCTGCGGCGCTTGGGACGGTGGCTGTCAAATACCAAGATATGATTGCGGCCCTGCCCGAAGCCACGCACGAGGCAATTGTAGGGGTCGGTAAACAGTGGTCGGGAGCGCGAGCACTTGAGGCGCTGCTGACTGTGGCGGGTGAGCTTAGGGGGCCTCCGCTCCAGCTCGACACCGGGCAGCTGCTGAAGATCGCGAAGAGAGGGGGAGTAACAGCGGTAGAGGCAGTGCACGCCTGGCGCAATGCGCTCACCGGGGCCCCCTTGAACCTGACCCCAGACCAGGTAGTCGCAATCGCGTCGCATGACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCAAATATCGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCAATGGGGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCCACGACGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGACTCACCCCAGACCAGGTAGTCGCAATCGCGTCGCATGACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCAAACGGAGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCCATGATGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCAATATTGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGGCTGACCCCAGACCAGGTAGTCGCAATCGCGTCGAATGGCGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCACATGACGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCCATGATGGAGGGAAACAGGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCAACGGTGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGTTTGACCCCAGACCAGGTAGTCGCAATCGCCAACAATAACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACGCCTGCACAAGTGGTCGCCATCGCCTCCAATATTGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGCCTGACACCCGAACAGGTGGTCGCCATTGCTAGCAACGGGGGAGGACGGCCAGCCTTGGAGTCCATCGTAGCCCAATTGTCCAGGCCCGATCCCGCGTTGGCTGCGTTAACGAATGACCATCTGGTGGCGTTGGCATGTCTTGGTGGACGACCCGCGCTCGATGCAGTCAAAAAGGGTCTGCCTCATGCTCCCGCATTGATCAAAAGAACCAACCGGCGGATTCCCGAGAGAACTTCCCATCGAGTCGCGGGATCCCAGCTGGTGAAGAGCGAGCTGGAGGAGAAGAAGTCCGAGCTGCGGCACAAGCTGAAGTACGTGCCCCACGAGTACATCGAGCTGATCGAGATCGCCAGGAACAGCACCCAGGACCGCATCCTGGAGATGAAGGTGATGGAGTTCTTCATGAAGGTGTACGGCTACAGGGGAAAGCACCTGGGCGGAAGCAGAAAGCCTGACGGCGCCATCTATACAGTGGGCAGCCCCATCGATTACGGCGTGATCGTGGACACAAAGGCCTACAGCGGCGGCTACAATCTGCCTATCGGCCAGGCCGACGAGATGCAGAGATACGTGGAGGAGAACCAGACCCGGAATAAGCACATCAACCCCAACGAGTGGTGGAAGGTGTACCCTAGCAGCGTGACCGAGTTCAAGTTCCTGTTCGTGAGCGGCCACTTCAAGGGCAACTACAAGGCCCAGCTGACCAGGCTGAACCACATCACCAACTGCAATGGCGCCGTGCTGAGCGTGGAGGAGCTGCTGATCGGCGGCGAGATGATCAAAGCCGGCACCCTGACACTGGAGGAGGTGCGGCGCAAGTTCAACAACGGCGAGATCAACTTCAGATCT
MDYKDHDGDYKDHDIDYKDDDDKMAPKKKRKVGIHRGVPMVDLRTLGYSQQQQEKIKPKVRSTVAQHHEALVGHGFTHAHIVALSQHPAALGTVAVKYQDMIAALPEATHEAIVGVGKQWSGARALEALLTVAGELRGPPLQLDTGQLLKIAKRGGVTAVEAVHAWRNALTGAPLNLTPDQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPAQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPDQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPAQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQDHGLTPDQVVAIANNNGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNGGGKQALETVQRLLPVLCQDHGLTPDQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPAQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGRPALESIVAQLSRPDPALAALTNDHLVALACLGGSQLVKSELEEKKSELRHKLKYVPHEYIELIEIARNSTQDRILEMKVMEFFMKVYGYRGKHLGGSRKPDGAIYTVGSPIDYGVIVDTKAYSGGYNLPIGQADEMQRYVEENQTRNKHINPNEWWKVYPSSVTEFKFLFVSGHFKGNYKAQLTRLNHITNCNGAVLSVEELLIGGEMIKAGTLTLEEVRRKFNNGEINFRS
ATGGACTACAAAGACCATGACGGTGATTATAAAGATCATGACATCGATTACAAGGATGACGATGACAAGATGGCCCCCAAGAAGAAGAGGAAGGTGGGCATTCACCGCGGGGTACCTATGGTGGACTTGAGGACACTCGGTTATTCGCAACAGCAACAGGAGAAAATCAAGCCTAAGGTCAGGAGCACCGTCGCGCAACACCACGAGGCGCTTGTGGGGCATGGCTTCACTCATGCGCATATTGTCGCGCTTTCACAGCACCCTGCGGCGCTTGGGACGGTGGCTGTCAAATACCAAGATATGATTGCGGCCCTGCCCGAAGCCACGCACGAGGCAATTGTAGGGGTCGGTAAACAGTGGTCGGGAGCGCGAGCACTTGAGGCGCTGCTGACTGTGGCGGGTGAGCTTAGGGGGCCTCCGCTCCAGCTCGACACCGGGCAGCTGCTGAAGATCGCGAAGAGAGGGGGAGTAACAGCGGTAGAGGCAGTGCACGCCTGGCGCAATGCGCTCACCGGGGCCCCCTTGAACCTGACCCCAGACCAGGTAGTCGCAATCGCGTCGCATGACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCACATGACGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCAATGGGGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCCACGACGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGACTCACCCCAGACCAGGTAGTCGCAATCGCGTCGAACATTGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCAAACGGAGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCCATGATGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCCACGACGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGGCTGACCCCAGACCAGGTAGTCGCAATCGCGTCGAATGGCGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGATCAAGTCGTGGCCATTGCAAATAATAACGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCAACATCGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCAACGGTGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGTTTGACCCCAGACCAGGTAGTCGCAATCGCGTCGAACATTGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCTAATATCGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCAACATCGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCCACGACGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGCCTGACACCCGAACAGGTGGTCGCCATTGCTAGCAACGGGGGAGGACGGCCAGCCTTGGAGTCCATCGTAGCCCAATTGTCCAGGCCCGATCCCGCGTTGGCTGCGTTAACGAATGACCATCTGGTGGCGTTGGCATGTCTTGGTGGATCCCAGCTGGTGAAGAGCGAGCTGGAGGAGAAGAAGTCCGAGCTGCGGCACAAGCTGAAGTACGTGCCCCACGAGTACATCGAGCTGATCGAGATCGCCAGGAACAGCACCCAGGACCGCATCCTGGAGATGAAGGTGATGGAGTTCTTCATGAAGGTGTACGGCTACAGGGGAAAGCACCTGGGCGGAAGCAGAAAGCCTGACGGCGCCATCTATACAGTGGGCAGCCCCATCGATTACGGCGTGATCGTGGACACAAAGGCCTACAGCGGCGGCTACAATCTGCCTATCGGCCAGGCCGACGAGATGCAGAGATACGTGGAGGAGAACCAGACCCGGAATAAGCACATCAACCCCAACGAGTGGTGGAAGGTGTACCCTAGCAGCGTGACCGAGTTCAAGTTCCTGTTCGTGAGCGGCCACTTCAAGGGCAACTACAAGGCCCAGCTGACCAGGCTGAACCACATCACCAACTGCAATGGCGCCGTGCTGAGCGTGGAGGAGCTGCTGATCGGCGGCGAGATGATCAAAGCCGGCACCCTGACACTGGAGGAGGTGCGGCGCAAGTTCAACAACGGCGAGATCAACTTCAGATCT
MDYKDHDGDYKDHDIDYKDDDDKMAPKKKRKVGIHRGVPMVDLRTLGYSQQQQEKIKPKVRSTVAQHHEALVGHGFTHAHIVALSQHPAALGTVAVKYQDMIAALPEATHEAIVGVGKQWSGARALEALLTVAGELRGPPLQLDTGQLLKIAKRGGVTAVEAVHAWRNALTGAPLNLTPDQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPAQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPDQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPAQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQDHGLTPDQVVAIANNNGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNGGGKQALETVQRLLPVLCQDHGLTPDQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPAQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGRPALESIVAQLSRPDPALAALTNDHLVALACLGGRPALDAVKKGLPHAPALIKRTNRRIPERTSHRVAGSQLVKSELEEKKSELRHKLKYVPHEYIELIEIARNSTQDRILEMKVMEFFMKVYGYRGKHLGGSRKPDGAIYTVGSPIDYGVIVDTKAYSGGYNLPIGQADEMQRYVEENQTRNKHINPNEWWKVYPSSVTEFKFLFVSGHFKGNYKAQLTRLNHITNCNGAVLSVEELLIGGEMIKAGTLTLEEVRRKFNNGEINFRS
ATGGACTACAAAGACCATGACGGTGATTATAAAGATCATGACATCGATTACAAGGATGACGATGACAAGATGGCCCCCAAGAAGAAGAGGAAGGTGGGCATTCACCGCGGGGTACCTATGGTGGACTTGAGGACACTCGGTTATTCGCAACAGCAACAGGAGAAAATCAAGCCTAAGGTCAGGAGCACCGTCGCGCAACACCACGAGGCGCTTGTGGGGCATGGCTTCACTCATGCGCATATTGTCGCGCTTTCACAGCACCCTGCGGCGCTTGGGACGGTGGCTGTCAAATACCAAGATATGATTGCGGCCCTGCCCGAAGCCACGCACGAGGCAATTGTAGGGGTCGGTAAACAGTGGTCGGGAGCGCGAGCACTTGAGGCGCTGCTGACTGTGGCGGGTGAGCTTAGGGGGCCTCCGCTCCAGCTCGACACCGGGCAGCTGCTGAAGATCGCGAAGAGAGGGGGAGTAACAGCGGTAGAGGCAGTGCACGCCTGGCGCAATGCGCTCACCGGGGCCCCCTTGAACCTGACCCCAGACCAGGTAGTCGCAATCGCGTCGCATGACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCACATGACGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCAATGGGGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCCACGACGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGACTCACCCCAGACCAGGTAGTCGCAATCGCGTCGAACATTGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCAAACGGAGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCCATGATGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCCACGACGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGGCTGACCCCAGACCAGGTAGTCGCAATCGCGTCGAATGGCGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGATCAAGTCGTGGCCATTGCAAATAATAACGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCAACATCGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCAACGGTGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGTTTGACCCCAGACCAGGTAGTCGCAATCGCGTCGAACATTGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCTAATATCGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCAACATCGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCCACGACGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGCCTGACACCCGAACAGGTGGTCGCCATTGCTAGCAACGGGGGAGGACGGCCAGCCTTGGAGTCCATCGTAGCCCAATTGTCCAGGCCCGATCCCGCGTTGGCTGCGTTAACGAATGACCATCTGGTGGCGTTGGCATGTCTTGGTGGACGACCCGCGCTCGATGCAGTCAAAAAGGGTCTGCCTCATGCTCCCGCATTGATCAAAAGAACCAACCGGCGGATTCCCGAGAGAACTTCCCATCGAGTCGCGGGATCCCAGCTGGTGAAGAGCGAGCTGGAGGAGAAGAAGTCCGAGCTGCGGCACAAGCTGAAGTACGTGCCCCACGAGTACATCGAGCTGATCGAGATCGCCAGGAACAGCACCCAGGACCGCATCCTGGAGATGAAGGTGATGGAGTTCTTCATGAAGGTGTACGGCTACAGGGGAAAGCACCTGGGCGGAAGCAGAAAGCCTGACGGCGCCATCTATACAGTGGGCAGCCCCATCGATTACGGCGTGATCGTGGACACAAAGGCCTACAGCGGCGGCTACAATCTGCCTATCGGCCAGGCCGACGAGATGCAGAGATACGTGGAGGAGAACCAGACCCGGAATAAGCACATCAACCCCAACGAGTGGTGGAAGGTGTACCCTAGCAGCGTGACCGAGTTCAAGTTCCTGTTCGTGAGCGGCCACTTCAAGGGCAACTACAAGGCCCAGCTGACCAGGCTGAACCACATCACCAACTGCAATGGCGCCGTGCTGAGCGTGGAGGAGCTGCTGATCGGCGGCGAGATGATCAAAGCCGGCACCCTGACACTGGAGGAGGTGCGGCGCAAGTTCAACAACGGCGAGATCAACTTCAGATCT
MDYKDHDGDYKDHDIDYKDDDDKMAPKKKRKVGIHRGVPMVDLRTLGYSQQQQEKIKPKVRSTVAQHHEALVGHGFTHAHIVALSQHPAALGTVAVKYQDMIAALPEATHEAIVGVGKQWSGARALEALLTVAGELRGPPLQLDTGQLLKIAKRGGVTAVEAVHAWRNALTGAPLNLTPDQVVAIANNNGGKQALETVQRLLPVLCQDHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPAQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPDQVVAIANNNGGKQALETVQRLLPVLCQDHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNGGGKQALETVQRLLPVLCQDHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQDHGLTPAQVVAIASNGGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGRPALESIVAQLSRPDPALAALTNDHLVALACLGGSQLVKSELEEKKSELRHKLKYVPHEYIELIEIARNSTQDRILEMKVMEFFMKVYGYRGKHLGGSRKPDGAIYTVGSPIDYGVIVDTKAYSGGYNLPIGQADEMQRYVEENQTRNKHINPNEWWKVYPSSVTEFKFLFVSGHFKGNYKAQLTRLNHITNCNGAVLSVEELLIGGEMIKAGTLTLEEVRRKFNNGEINFRS
ATGGACTACAAAGACCATGACGGTGATTATAAAGATCATGACATCGATTACAAGGATGACGATGACAAGATGGCCCCCAAGAAGAAGAGGAAGGTGGGCATTCACCGCGGGGTACCTATGGTGGACTTGAGGACACTCGGTTATTCGCAACAGCAACAGGAGAAAATCAAGCCTAAGGTCAGGAGCACCGTCGCGCAACACCACGAGGCGCTTGTGGGGCATGGCTTCACTCATGCGCATATTGTCGCGCTTTCACAGCACCCTGCGGCGCTTGGGACGGTGGCTGTCAAATACCAAGATATGATTGCGGCCCTGCCCGAAGCCACGCACGAGGCAATTGTAGGGGTCGGTAAACAGTGGTCGGGAGCGCGAGCACTTGAGGCGCTGCTGACTGTGGCGGGTGAGCTTAGGGGGCCTCCGCTCCAGCTCGACACCGGGCAGCTGCTGAAGATCGCGAAGAGAGGGGGAGTAACAGCGGTAGAGGCAGTGCACGCCTGGCGCAATGCGCTCACCGGGGCCCCCTTGAACCTGACCCCAGACCAGGTAGTCGCAATCGCCAACAATAACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCACATGACGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCAATGGGGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCCACGACGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGACTCACCCCAGACCAGGTAGTCGCAATCGCGTCGAATGGCGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCAAACGGAGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCCATGATGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCAATATTGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGGCTGACCCCAGACCAGGTAGTCGCAATCGCCAACAATAACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCACATGACGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCCATGATGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCAACGGTGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGTTTGACCCCAGACCAGGTAGTCGCAATCGCGTCGAATGGCGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACGCCTGCACAAGTGGTCGCCATCGCCTCCAACGGTGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGCCTGACACCCGAACAGGTGGTCGCCATTGCTAGCAACGGGGGAGGACGGCCAGCCTTGGAGTCCATCGTAGCCCAATTGTCCAGGCCCGATCCCGCGTTGGCTGCGTTAACGAATGACCATCTGGTGGCGTTGGCATGTCTTGGTGGATCCCAGCTGGTGAAGAGCGAGCTGGAGGAGAAGAAGTCCGAGCTGCGGCACAAGCTGAAGTACGTGCCCCACGAGTACATCGAGCTGATCGAGATCGCCAGGAACAGCACCCAGGACCGCATCCTGGAGATGAAGGTGATGGAGTTCTTCATGAAGGTGTACGGCTACAGGGGAAAGCACCTGGGCGGAAGCAGAAAGCCTGACGGCGCCATCTATACAGTGGGCAGCCCCATCGATTACGGCGTGATCGTGGACACAAAGGCCTACAGCGGCGGCTACAATCTGCCTATCGGCCAGGCCGACGAGATGCAGAGATACGTGGAGGAGAACCAGACCCGGAATAAGCACATCAACCCCAACGAGTGGTGGAAGGTGTACCCTAGCAGCGTGACCGAGTTCAAGTTCCTGTTCGTGAGCGGCCACTTCAAGGGCAACTACAAGGCCCAGCTGACCAGGCTGAACCACATCACCAACTGCAATGGCGCCGTGCTGAGCGTGGAGGAGCTGCTGATCGGCGGCGAGATGATCAAAGCCGGCACCCTGACACTGGAGGAGGTGCGGCGCAAGTTCAACAACGGCGAGATCAACTTCAGATCT
MDYKDHDGDYKDHDIDYKDDDDKMAPKKKRKVGIHRGVPMVDLRTLGYSQQQQEKIKPKVRSTVAQHHEALVGHGFTHAHIVALSQHPAALGTVAVKYQDMIAALPEATHEAIVGVGKQWSGARALEALLTVAGELRGPPLQLDTGQLLKIAKRGGVTAVEAVHAWRNALTGAPLNLTPDQVVAIANNNGGKQALETVQRLLPVLCQDHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPAQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPDQVVAIANNNGGKQALETVQRLLPVLCQDHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNGGGKQALETVQRLLPVLCQDHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQDHGLTPAQVVAIASNGGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGRPALESIVAQLSRPDPALAALTNDHLVALACLGGRPALDAVKKGLPHAPALIKRTNRRIPERTSHRVAGSQLVKSELEEKKSELRHKLKYVPHEYIELIEIARNSTQDRILEMKVMEFFMKVYGYRGKHLGGSRKPDGAIYTVGSPIDYGVIVDTKAYSGGYNLPIGQADEMQRYVEENQTRNKHINPNEWWKVYPSSVTEFKFLFVSGHFKGNYKAQLTRLNHITNCNGAVLSVEELLIGGEMIKAGTLTLEEVRRKFNNGEINFRS
ATGGACTACAAAGACCATGACGGTGATTATAAAGATCATGACATCGATTACAAGGATGACGATGACAAGATGGCCCCCAAGAAGAAGAGGAAGGTGGGCATTCACCGCGGGGTACCTATGGTGGACTTGAGGACACTCGGTTATTCGCAACAGCAACAGGAGAAAATCAAGCCTAAGGTCAGGAGCACCGTCGCGCAACACCACGAGGCGCTTGTGGGGCATGGCTTCACTCATGCGCATATTGTCGCGCTTTCACAGCACCCTGCGGCGCTTGGGACGGTGGCTGTCAAATACCAAGATATGATTGCGGCCCTGCCCGAAGCCACGCACGAGGCAATTGTAGGGGTCGGTAAACAGTGGTCGGGAGCGCGAGCACTTGAGGCGCTGCTGACTGTGGCGGGTGAGCTTAGGGGGCCTCCGCTCCAGCTCGACACCGGGCAGCTGCTGAAGATCGCGAAGAGAGGGGGAGTAACAGCGGTAGAGGCAGTGCACGCCTGGCGCAATGCGCTCACCGGGGCCCCCTTGAACCTGACCCCAGACCAGGTAGTCGCAATCGCCAACAATAACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCACATGACGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCAATGGGGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCCACGACGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGACTCACCCCAGACCAGGTAGTCGCAATCGCGTCGAATGGCGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCAAACGGAGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCCATGATGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCAATATTGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGGCTGACCCCAGACCAGGTAGTCGCAATCGCCAACAATAACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCACATGACGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCCATGATGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCAACGGTGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGTTTGACCCCAGACCAGGTAGTCGCAATCGCGTCGAATGGCGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACGCCTGCACAAGTGGTCGCCATCGCCTCCAACGGTGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGCTTGACACCCGAACAGGTGGTCGCCATTGCTAGCAACGGGGGAGGACGGCCAGCCTTGGAGTCCATCGTAGCCCAATTGTCCAGGCCCGATCCCGCGTTGGCTGCGTTAACGAATGACCATCTGGTGGCGTTGGCATGTCTTGGTGGACGACCCGCGCTCGATGCAGTCAAAAAGGGTCTGCCTCATGCTCCCGCATTGATCAAAAGAACCAACCGGCGGATTCCCGAGAGAACTTCCCATCGAGTCGCGGGATCCCAGCTGGTGAAGAGCGAGCTGGAGGAGAAGAAGTCCGAGCTGCGGCACAAGCTGAAGTACGTGCCCCACGAGTACATCGAGCTGATCGAGATCGCCAGGAACAGCACCCAGGACCGCATCCTGGAGATGAAGGTGATGGAGTTCTTCATGAAGGTGTACGGCTACAGGGGAAAGCACCTGGGCGGAAGCAGAAAGCCTGACGGCGCCATCTATACAGTGGGCAGCCCCATCGATTACGGCGTGATCGTGGACACAAAGGCCTACAGCGGCGGCTACAATCTGCCTATCGGCCAGGCCGACGAGATGCAGAGATACGTGGAGGAGAACCAGACCCGGAATAAGCACATCAACCCCAACGAGTGGTGGAAGGTGTACCCTAGCAGCGTGACCGAGTTCAAGTTCCTGTTCGTGAGCGGCCACTTCAAGGGCAACTACAAGGCCCAGCTGACCAGGCTGAACCACATCACCAACTGCAATGGCGCCGTGCTGAGCGTGGAGGAGCTGCTGATCGGCGGCGAGATGATCAAAGCCGGCACCCTGACACTGGAGGAGGTGCGGCGCAAGTTCAACAACGGCGAGATCAACTTCAGATCT
MDYKDHDGDYKDHDIDYKDDDDKMAPKKKRKVGIHRGVPMVDLRTLGYSQQQQEKIKPKVRSTVAQHHEALVGHGFTHAHIVALSQHPAALGTVAVKYQDMIAALPEATHEAIVGVGKQWSGARALEALLTVAGELRGPPLQLDTGQLLKIAKRGGVTAVEAVHAWRNALTGAPLNLTPDQVVAIANNNGGKQALETVQRLLPVLCQDHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPAQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPDQVVAIANNNGGKQALETVQRLLPVLCQDHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNGGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGRPALESIVAQLSRPDPALAALTNDHLVALACLGGSQLVKSELEEKKSELRHKLKYVPHEYIELIEIARNSTQDRILEMKVMEFFMKVYGYRGKHLGGSRKPDGAIYTVGSPIDYGVIVDTKAYSGGYNLPIGQADEMQRYVEENQTRNKHINPNEWWKVYPSSVTEFKFLFVSGHFKGNYKAQLTRLNHITNCNGAVLSVEELLIGGEMIKAGTLTLEEVRRKFNNGEINFRS
ATGGACTACAAAGACCATGACGGTGATTATAAAGATCATGACATCGATTACAAGGATGACGATGACAAGATGGCCCCCAAGAAGAAGAGGAAGGTGGGCATTCACCGCGGGGTACCTATGGTGGACTTGAGGACACTCGGTTATTCGCAACAGCAACAGGAGAAAATCAAGCCTAAGGTCAGGAGCACCGTCGCGCAACACCACGAGGCGCTTGTGGGGCATGGCTTCACTCATGCGCATATTGTCGCGCTTTCACAGCACCCTGCGGCGCTTGGGACGGTGGCTGTCAAATACCAAGATATGATTGCGGCCCTGCCCGAAGCCACGCACGAGGCAATTGTAGGGGTCGGTAAACAGTGGTCGGGAGCGCGAGCACTTGAGGCGCTGCTGACTGTGGCGGGTGAGCTTAGGGGGCCTCCGCTCCAGCTCGACACCGGGCAGCTGCTGAAGATCGCGAAGAGAGGGGGAGTAACAGCGGTAGAGGCAGTGCACGCCTGGCGCAATGCGCTCACCGGGGCCCCCTTGAACCTGACCCCAGACCAGGTAGTCGCAATCGCCAACAATAACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCACATGACGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCAATGGGGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCCACGACGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGACTCACCCCAGACCAGGTAGTCGCAATCGCGTCGAATGGCGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCAAACGGAGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCCATGATGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCAATATTGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGGCTGACCCCAGACCAGGTAGTCGCAATCGCCAACAATAACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCACATGACGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCCATGATGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCAACGGTGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGCCTGACACCCGAACAGGTGGTCGCCATTGCTAGCAACGGGGGAGGACGGCCAGCCTTGGAGTCCATCGTAGCCCAATTGTCCAGGCCCGATCCCGCGTTGGCTGCGTTAACGAATGACCATCTGGTGGCGTTGGCATGTCTTGGTGGATCCCAGCTGGTGAAGAGCGAGCTGGAGGAGAAGAAGTCCGAGCTGCGGCACAAGCTGAAGTACGTGCCCCACGAGTACATCGAGCTGATCGAGATCGCCAGGAACAGCACCCAGGACCGCATCCTGGAGATGAAGGTGATGGAGTTCTTCATGAAGGTGTACGGCTACAGGGGAAAGCACCTGGGCGGAAGCAGAAAGCCTGACGGCGCCATCTATACAGTGGGCAGCCCCATCGATTACGGCGTGATCGTGGACACAAAGGCCTACAGCGGCGGCTACAATCTGCCTATCGGCCAGGCCGACGAGATGCAGAGATACGTGGAGGAGAACCAGACCCGGAATAAGCACATCAACCCCAACGAGTGGTGGAAGGTGTACCCTAGCAGCGTGACCGAGTTCAAGTTCCTGTTCGTGAGCGGCCACTTCAAGGGCAACTACAAGGCCCAGCTGACCAGGCTGAACCACATCACCAACTGCAATGGCGCCGTGCTGAGCGTGGAGGAGCTGCTGATCGGCGGCGAGATGATCAAAGCCGGCACCCTGACACTGGAGGAGGTGCGGCGCAAGTTCAACAACGGCGAGATCAACTTCAGATCT
MDYKDHDGDYKDHDIDYKDDDDKMAPKKKRKVGIHRGVPMVDLRTLGYSQQQQEKIKPKVRSTVAQHHEALVGHGFTHAHIVALSQHPAALGTVAVKYQDMIAALPEATHEAIVGVGKQWSGARALEALLTVAGELRGPPLQLDTGQLLKIAKRGGVTAVEAVHAWRNALTGAPLNLTPDQVVAIANNNGGKQALETVQRLLPVLCQDHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPAQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPDQVVAIANNNGGKQALETVQRLLPVLCQDHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNGGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGRPALESIVAQLSRPDPALAALTNDHLVALACLGGRPALDAVKKGLPHAPALIKRTNRRIPERTSHRVAGSQLVKSELEEKKSELRHKLKYVPHEYIELIEIARNSTQDRILEMKVMEFFMKVYGYRGKHLGGSRKPDGAIYTVGSPIDYGVIVDTKAYSGGYNLPIGQADEMQRYVEENQTRNKHINPNEWWKVYPSSVTEFKFLFVSGHFKGNYKAQLTRLNHITNCNGAVLSVEELLIGGEMIKAGTLTLEEVRRKFNNGEINFRS
ATGGACTACAAAGACCATGACGGTGATTATAAAGATCATGACATCGATTACAAGGATGACGATGACAAGATGGCCCCCAAGAAGAAGAGGAAGGTGGGCATTCACCGCGGGGTACCTATGGTGGACTTGAGGACACTCGGTTATTCGCAACAGCAACAGGAGAAAATCAAGCCTAAGGTCAGGAGCACCGTCGCGCAACACCACGAGGCGCTTGTGGGGCATGGCTTCACTCATGCGCATATTGTCGCGCTTTCACAGCACCCTGCGGCGCTTGGGACGGTGGCTGTCAAATACCAAGATATGATTGCGGCCCTGCCCGAAGCCACGCACGAGGCAATTGTAGGGGTCGGTAAACAGTGGTCGGGAGCGCGAGCACTTGAGGCGCTGCTGACTGTGGCGGGTGAGCTTAGGGGGCCTCCGCTCCAGCTCGACACCGGGCAGCTGCTGAAGATCGCGAAGAGAGGGGGAGTAACAGCGGTAGAGGCAGTGCACGCCTGGCGCAATGCGCTCACCGGGGCCCCCTTGAACCTGACCCCAGACCAGGTAGTCGCAATCGCCAACAATAACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCACATGACGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCAATGGGGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCCACGACGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGACTCACCCCAGACCAGGTAGTCGCAATCGCGTCGAATGGCGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCAAACGGAGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCCATGATGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCAATATTGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGGCTGACCCCAGACCAGGTAGTCGCAATCGCCAACAATAACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCACATGACGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCCATGATGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCAACGGTGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGCCTGACACCCGAACAGGTGGTCGCCATTGCTAGCAACGGGGGAGGACGGCCAGCCTTGGAGTCCATCGTAGCCCAATTGTCCAGGCCCGATCCCGCGTTGGCTGCGTTAACGAATGACCATCTGGTGGCGTTGGCATGTCTTGGTGGACGACCCGCGCTCGATGCAGTCAAAAAGGGTCTGCCTCATGCTCCCGCATTGATCAAAAGAACCAACCGGCGGATTCCCGAGAGAACTTCCCATCGAGTCGCGGGATCCCAGCTGGTGAAGAGCGAGCTGGAGGAGAAGAAGTCCGAGCTGCGGCACAAGCTGAAGTACGTGCCCCACGAGTACATCGAGCTGATCGAGATCGCCAGGAACAGCACCCAGGACCGCATCCTGGAGATGAAGGTGATGGAGTTCTTCATGAAGGTGTACGGCTACAGGGGAAAGCACCTGGGCGGAAGCAGAAAGCCTGACGGCGCCATCTATACAGTGGGCAGCCCCATCGATTACGGCGTGATCGTGGACACAAAGGCCTACAGCGGCGGCTACAATCTGCCTATCGGCCAGGCCGACGAGATGCAGAGATACGTGGAGGAGAACCAGACCCGGAATAAGCACATCAACCCCAACGAGTGGTGGAAGGTGTACCCTAGCAGCGTGACCGAGTTCAAGTTCCTGTTCGTGAGCGGCCACTTCAAGGGCAACTACAAGGCCCAGCTGACCAGGCTGAACCACATCACCAACTGCAATGGCGCCGTGCTGAGCGTGGAGGAGCTGCTGATCGGCGGCGAGATGATCAAAGCCGGCACCCTGACACTGGAGGAGGTGCGGCGCAAGTTCAACAACGGCGAGATCAACTTCAGATCT
MDYKDHDGDYKDHDIDYKDDDDKMAPKKKRKVGIHRGVPMVDLRTLGYSQQQQEKIKPKVRSTVAQHHEALVGHGFTHAHIVALSQHPAALGTVAVKYQDMIAALPEATHEAIVGVGKQWSGARALEALLTVAGELRGPPLQLDTGQLLKIAKRGGVTAVEAVHAWRNALTGAPLNLTPDQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPAQVVAIANNNGGKQALETVQRLLPVLCQDHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNGGGKQALETVQRLLPVLCQDHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQDHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPAQVVAIANNNGGKQALETVQRLLPVLCQDHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPAQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGRPALESIVAQLSRPDPALAALTNDHLVALACLGGSQLVKSELEEKKSELRHKLKYVPHEYIELIEIARNSTQDRILEMKVMEFFMKVYGYRGKHLGGSRKPDGAIYTVGSPIDYGVIVDTKAYSGGYNLPIGQADEMQRYVEENQTRNKHINPNEWWKVYPSSVTEFKFLFVSGHFKGNYKAQLTRLNHITNCNGAVLSVEELLIGGEMIKAGTLTLEEVRRKFNNGEINFRS
ATGGACTACAAAGACCATGACGGTGATTATAAAGATCATGACATCGATTACAAGGATGACGATGACAAGATGGCCCCCAAGAAGAAGAGGAAGGTGGGCATTCACCGCGGGGTACCTATGGTGGACTTGAGGACACTCGGTTATTCGCAACAGCAACAGGAGAAAATCAAGCCTAAGGTCAGGAGCACCGTCGCGCAACACCACGAGGCGCTTGTGGGGCATGGCTTCACTCATGCGCATATTGTCGCGCTTTCACAGCACCCTGCGGCGCTTGGGACGGTGGCTGTCAAATACCAAGATATGATTGCGGCCCTGCCCGAAGCCACGCACGAGGCAATTGTAGGGGTCGGTAAACAGTGGTCGGGAGCGCGAGCACTTGAGGCGCTGCTGACTGTGGCGGGTGAGCTTAGGGGGCCTCCGCTCCAGCTCGACACCGGGCAGCTGCTGAAGATCGCGAAGAGAGGGGGAGTAACAGCGGTAGAGGCAGTGCACGCCTGGCGCAATGCGCTCACCGGGGCCCCCTTGAACCTGACCCCAGACCAGGTAGTCGCAATCGCGTCGCATGACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCAAATATCGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCAATGGGGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCAACAACAACGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGACTCACCCCAGACCAGGTAGTCGCAATCGCGTCGCATGACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCAAACGGAGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCCATGATGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCAACGGTGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGGCTGACCCCAGACCAGGTAGTCGCAATCGCGTCGAATGGCGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCACATGACGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCAACATCGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCAACAACAACGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGTTTGACCCCAGACCAGGTAGTCGCAATCGCGTCGCATGACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACGCCTGCACAAGTGGTCGCCATCGCCTCCCACGACGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGCCTGACACCCGAACAGGTGGTCGCCATTGCTAGCAACGGGGGAGGACGGCCAGCCTTGGAGTCCATCGTAGCCCAATTGTCCAGGCCCGATCCCGCGTTGGCTGCGTTAACGAATGACCATCTGGTGGCGTTGGCATGTCTTGGTGGATCCCAGCTGGTGAAGAGCGAGCTGGAGGAGAAGAAGTCCGAGCTGCGGCACAAGCTGAAGTACGTGCCCCACGAGTACATCGAGCTGATCGAGATCGCCAGGAACAGCACCCAGGACCGCATCCTGGAGATGAAGGTGATGGAGTTCTTCATGAAGGTGTACGGCTACAGGGGAAAGCACCTGGGCGGAAGCAGAAAGCCTGACGGCGCCATCTATACAGTGGGCAGCCCCATCGATTACGGCGTGATCGTGGACACAAAGGCCTACAGCGGCGGCTACAATCTGCCTATCGGCCAGGCCGACGAGATGCAGAGATACGTGGAGGAGAACCAGACCCGGAATAAGCACATCAACCCCAACGAGTGGTGGAAGGTGTACCCTAGCAGCGTGACCGAGTTCAAGTTCCTGTTCGTGAGCGGCCACTTCAAGGGCAACTACAAGGCCCAGCTGACCAGGCTGAACCACATCACCAACTGCAATGGCGCCGTGCTGAGCGTGGAGGAGCTGCTGATCGGCGGCGAGATGATCAAAGCCGGCACCCTGACACTGGAGGAGGTGCGGCGCAAGTTCAACAACGGCGAGATCAACTTCAGATCT
MDYKDHDGDYKDHDIDYKDDDDKMAPKKKRKVGIHRGVPMVDLRTLGYSQQQQEKIKPKVRSTVAQHHEALVGHGFTHAHIVALSQHPAALGTVAVKYQDMIAALPEATHEAIVGVGKQWSGARALEALLTVAGELRGPPLQLDTGQLLKIAKRGGVTAVEAVHAWRNALTGAPLNLTPDQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPAQVVAIANNNGGKQALETVQRLLPVLCQDHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNGGGKQALETVQRLLPVLCQDHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQDHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPAQVVAIANNNGGKQALETVQRLLPVLCQDHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPAQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGRPALESIVAQLSRPDPALAALTNDHLVALACLGGRPALDAVKKGLPHAPALIKRTNRRIPERTSHRVAGSQLVKSELEEKKSELRHKLKYVPHEYIELIEIARNSTQDRILEMKVMEFFMKVYGYRGKHLGGSRKPDGAIYTVGSPIDYGVIVDTKAYSGGYNLPIGQADEMQRYVEENQTRNKHINPNEWWKVYPSSVTEFKFLFVSGHFKGNYKAQLTRLNHITNCNGAVLSVEELLIGGEMIKAGTLTLEEVRRKFNNGEINFRS
ATGGACTACAAAGACCATGACGGTGATTATAAAGATCATGACATCGATTACAAGGATGACGATGACAAGATGGCCCCCAAGAAGAAGAGGAAGGTGGGCATTCACCGCGGGGTACCTATGGTGGACTTGAGGACACTCGGTTATTCGCAACAGCAACAGGAGAAAATCAAGCCTAAGGTCAGGAGCACCGTCGCGCAACACCACGAGGCGCTTGTGGGGCATGGCTTCACTCATGCGCATATTGTCGCGCTTTCACAGCACCCTGCGGCGCTTGGGACGGTGGCTGTCAAATACCAAGATATGATTGCGGCCCTGCCCGAAGCCACGCACGAGGCAATTGTAGGGGTCGGTAAACAGTGGTCGGGAGCGCGAGCACTTGAGGCGCTGCTGACTGTGGCGGGTGAGCTTAGGGGGCCTCCGCTCCAGCTCGACACCGGGCAGCTGCTGAAGATCGCGAAGAGAGGGGGAGTAACAGCGGTAGAGGCAGTGCACGCCTGGCGCAATGCGCTCACCGGGGCCCCCTTGAACCTGACCCCAGACCAGGTAGTCGCAATCGCGTCGCATGACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCAAATATCGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCAATGGGGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCAACAACAACGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGACTCACCCCAGACCAGGTAGTCGCAATCGCGTCGCATGACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCAAACGGAGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCCATGATGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCAACGGTGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGGCTGACCCCAGACCAGGTAGTCGCAATCGCGTCGAATGGCGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCACATGACGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCAACATCGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCAACAACAACGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGTTTGACCCCAGACCAGGTAGTCGCAATCGCGTCGCATGACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACGCCTGCACAAGTGGTCGCCATCGCCTCCCACGACGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGCCTGACACCCGAACAGGTGGTCGCCATTGCTAGCAACGGGGGAGGACGGCCAGCCTTGGAGTCCATCGTAGCCCAATTGTCCAGGCCCGATCCCGCGTTGGCTGCGTTAACGAATGACCATCTGGTGGCGTTGGCATGTCTTGGTGGACGACCCGCGCTCGATGCAGTCAAAAAGGGTCTGCCTCATGCTCCCGCATTGATCAAAAGAACCAACCGGCGGATTCCCGAGAGAACTTCCCATCGAGTCGCGGGATCCCAGCTGGTGAAGAGCGAGCTGGAGGAGAAGAAGTCCGAGCTGCGGCACAAGCTGAAGTACGTGCCCCACGAGTACATCGAGCTGATCGAGATCGCCAGGAACAGCACCCAGGACCGCATCCTGGAGATGAAGGTGATGGAGTTCTTCATGAAGGTGTACGGCTACAGGGGAAAGCACCTGGGCGGAAGCAGAAAGCCTGACGGCGCCATCTATACAGTGGGCAGCCCCATCGATTACGGCGTGATCGTGGACACAAAGGCCTACAGCGGCGGCTACAATCTGCCTATCGGCCAGGCCGACGAGATGCAGAGATACGTGGAGGAGAACCAGACCCGGAATAAGCACATCAACCCCAACGAGTGGTGGAAGGTGTACCCTAGCAGCGTGACCGAGTTCAAGTTCCTGTTCGTGAGCGGCCACTTCAAGGGCAACTACAAGGCCCAGCTGACCAGGCTGAACCACATCACCAACTGCAATGGCGCCGTGCTGAGCGTGGAGGAGCTGCTGATCGGCGGCGAGATGATCAAAGCCGGCACCCTGACACTGGAGGAGGTGCGGCGCAAGTTCAACAACGGCGAGATCAACTTCAGATCT
MDYKDHDGDYKDHDIDYKDDDDKMAPKKKRKVGIHRGVPMVDLRTLGYSQQQQEKIKPKVRSTVAQHHEALVGHGFTHAHIVALSQHPAALGTVAVKYQDMIAALPEATHEAIVGVGKQWSGARALEALLTVAGELRGPPLQLDTGQLLKIAKRGGVTAVEAVHAWRNALTGAPLNLTPDQVVAIANNNGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPDQVVAIANNNGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQDHGLTPDQVVAIANNNGGKQALETVQRLLPVLCQAHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNGGGKQALETVQRLLPVLCQDHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGRPALESIVAQLSRPDPALAALTNDHLVALACLGGSQLVKSELEEKKSELRHKLKYVPHEYIELIEIARNSTQDRILEMKVMEFFMKVYGYRGKHLGGSRKPDGAIYTVGSPIDYGVIVDTKAYSGGYNLPIGQADEMQRYVEENQTRNKHINPNEWWKVYPSSVTEFKFLFVSGHFKGNYKAQLTRLNHITNCNGAVLSVEELLIGGEMIKAGTLTLEEVRRKFNNGEINFRS
ATGGACTACAAAGACCATGACGGTGATTATAAAGATCATGACATCGATTACAAGGATGACGATGACAAGATGGCCCCCAAGAAGAAGAGGAAGGTGGGCATTCACCGCGGGGTACCTATGGTGGACTTGAGGACACTCGGTTATTCGCAACAGCAACAGGAGAAAATCAAGCCTAAGGTCAGGAGCACCGTCGCGCAACACCACGAGGCGCTTGTGGGGCATGGCTTCACTCATGCGCATATTGTCGCGCTTTCACAGCACCCTGCGGCGCTTGGGACGGTGGCTGTCAAATACCAAGATATGATTGCGGCCCTGCCCGAAGCCACGCACGAGGCAATTGTAGGGGTCGGTAAACAGTGGTCGGGAGCGCGAGCACTTGAGGCGCTGCTGACTGTGGCGGGTGAGCTTAGGGGGCCTCCGCTCCAGCTCGACACCGGGCAGCTGCTGAAGATCGCGAAGAGAGGGGGAGTAACAGCGGTAGAGGCAGTGCACGCCTGGCGCAATGCGCTCACCGGGGCCCCCTTGAACCTGACCCCAGACCAGGTAGTCGCAATCGCCAACAATAACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCAAACGGAGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCCATGATGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCAATATTGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGACTCACCCCAGACCAGGTAGTCGCAATCGCCAACAATAACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCAAACGGAGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCCATGATGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCAATATTGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGGCTGACCCCAGACCAGGTAGTCGCAATCGCGTCGAATGGCGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGATCAAGTCGTGGCCATTGCAAATAATAACGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCCATGATGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCAACGGTGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGTTTGACCCCAGACCAGGTAGTCGCAATCGCGTCGCATGACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTACTGTGCCAGGATCATGGCCTGACACCCGAACAGGTGGTCGCCATTGCTAGCAACGGGGGAGGACGGCCAGCCTTGGAGTCCATCGTAGCCCAATTGTCCAGGCCCGATCCCGCGTTGGCTGCGTTAACGAATGACCATCTGGTGGCGTTGGCATGTCTTGGTGGATCCCAGCTGGTGAAGAGCGAGCTGGAGGAGAAGAAGTCCGAGCTGCGGCACAAGCTGAAGTACGTGCCCCACGAGTACATCGAGCTGATCGAGATCGCCAGGAACAGCACCCAGGACCGCATCCTGGAGATGAAGGTGATGGAGTTCTTCATGAAGGTGTACGGCTACAGGGGAAAGCACCTGGGCGGAAGCAGAAAGCCTGACGGCGCCATCTATACAGTGGGCAGCCCCATCGATTACGGCGTGATCGTGGACACAAAGGCCTACAGCGGCGGCTACAATCTGCCTATCGGCCAGGCCGACGAGATGCAGAGATACGTGGAGGAGAACCAGACCCGGAATAAGCACATCAACCCCAACGAGTGGTGGAAGGTGTACCCTAGCAGCGTGACCGAGTTCAAGTTCCTGTTCGTGAGCGGCCACTTCAAGGGCAACTACAAGGCCCAGCTGACCAGGCTGAACCACATCACCAACTGCAATGGCGCCGTGCTGAGCGTGGAGGAGCTGCTGATCGGCGGCGAGATGATCAAAGCCGGCACCCTGACACTGGAGGAGGTGCGGCGCAAGTTCAACAACGGCGAGATCAACTTCAGATCT
MDYKDHDGDYKDHDIDYKDDDDKMAPKKKRKVGIHRGVPMVDLRTLGYSQQQQEKIKPKVRSTVAQHHEALVGHGFTHAHIVALSQHPAALGTVAVKYQDMIAALPEATHEAIVGVGKQWSGARALEALLTVAGELRGPPLQLDTGQLLKIAKRGGVTAVEAVHAWRNALTGAPLNLTPDQVVAIANNNGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPDQVVAIANNNGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQDHGLTPDQVVAIANNNGGKQALETVQRLLPVLCQAHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNGGGKQALETVQRLLPVLCQDHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGRPALESIVAQLSRPDPALAALTNDHLVALACLGGRPALDAVKKGLPHAPALIKRTNRRIPERTSHRVAGSQLVKSELEEKKSELRHKLKYVPHEYIELIEIARNSTQDRILEMKVMEFFMKVYGYRGKHLGGSRKPDGAIYTVGSPIDYGVIVDTKAYSGGYNLPIGQADEMQRYVEENQTRNKHINPNEWWKVYPSSVTEFKFLFVSGHFKGNYKAQLTRLNHITNCNGAVLSVEELLIGGEMIKAGTLTLEEVRRKFNNGEINFRS
ATGGACTACAAAGACCATGACGGTGATTATAAAGATCATGACATCGATTACAAGGATGACGATGACAAGATGGCCCCCAAGAAGAAGAGGAAGGTGGGCATTCACCGCGGGGTACCTATGGTGGACTTGAGGACACTCGGTTATTCGCAACAGCAACAGGAGAAAATCAAGCCTAAGGTCAGGAGCACCGTCGCGCAACACCACGAGGCGCTTGTGGGGCATGGCTTCACTCATGCGCATATTGTCGCGCTTTCACAGCACCCTGCGGCGCTTGGGACGGTGGCTGTCAAATACCAAGATATGATTGCGGCCCTGCCCGAAGCCACGCACGAGGCAATTGTAGGGGTCGGTAAACAGTGGTCGGGAGCGCGAGCACTTGAGGCGCTGCTGACTGTGGCGGGTGAGCTTAGGGGGCCTCCGCTCCAGCTCGACACCGGGCAGCTGCTGAAGATCGCGAAGAGAGGGGGAGTAACAGCGGTAGAGGCAGTGCACGCCTGGCGCAATGCGCTCACCGGGGCCCCCTTGAACCTGACCCCAGACCAGGTAGTCGCAATCGCCAACAATAACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCAAACGGAGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCCATGATGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCAATATTGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGACTCACCCCAGACCAGGTAGTCGCAATCGCCAACAATAACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCAAACGGAGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCCATGATGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCAATATTGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGGCTGACCCCAGACCAGGTAGTCGCAATCGCGTCGAATGGCGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGATCAAGTCGTGGCCATTGCAAATAATAACGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCCATGATGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCAACGGTGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGTTTGACCCCAGACCAGGTAGTCGCAATCGCGTCGCATGACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTACTGTGCCAGGATCATGGCCTGACACCCGAACAGGTGGTCGCCATTGCTAGCAACGGGGGAGGACGGCCAGCCTTGGAGTCCATCGTAGCCCAATTGTCCAGGCCCGATCCCGCGTTGGCTGCGTTAACGAATGACCATCTGGTGGCGTTGGCATGTCTTGGTGGACGACCCGCGCTCGATGCAGTCAAAAAGGGTCTGCCTCATGCTCCCGCATTGATCAAAAGAACCAACCGGCGGATTCCCGAGAGAACTTCCCATCGAGTCGCGGGATCCCAGCTGGTGAAGAGCGAGCTGGAGGAGAAGAAGTCCGAGCTGCGGCACAAGCTGAAGTACGTGCCCCACGAGTACATCGAGCTGATCGAGATCGCCAGGAACAGCACCCAGGACCGCATCCTGGAGATGAAGGTGATGGAGTTCTTCATGAAGGTGTACGGCTACAGGGGAAAGCACCTGGGCGGAAGCAGAAAGCCTGACGGCGCCATCTATACAGTGGGCAGCCCCATCGATTACGGCGTGATCGTGGACACAAAGGCCTACAGCGGCGGCTACAATCTGCCTATCGGCCAGGCCGACGAGATGCAGAGATACGTGGAGGAGAACCAGACCCGGAATAAGCACATCAACCCCAACGAGTGGTGGAAGGTGTACCCTAGCAGCGTGACCGAGTTCAAGTTCCTGTTCGTGAGCGGCCACTTCAAGGGCAACTACAAGGCCCAGCTGACCAGGCTGAACCACATCACCAACTGCAATGGCGCCGTGCTGAGCGTGGAGGAGCTGCTGATCGGCGGCGAGATGATCAAAGCCGGCACCCTGACACTGGAGGAGGTGCGGCGCAAGTTCAACAACGGCGAGATCAACTTCAGATCT
MDYKDHDGDYKDHDIDYKDDDDKMAPKKKRKVGIHRGVPMVDLRTLGYSQQQQEKIKPKVRSTVAQHHEALVGHGFTHAHIVALSQHPAALGTVAVKYQDMIAALPEATHEAIVGVGKQWSGARALEALLTVAGELRGPPLQLDTGQLLKIAKRGGVTAVEAVHAWRNALTGAPLNLTPDQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPAQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPDQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPAQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGRPALESIVAQLSRPDPALAALTNDHLVALACLGGSQLVKSELEEKKSELRHKLKYVPHEYIELIEIARNSTQDRILEMKVMEFFMKVYGYRGKHLGGSRKPDGAIYTVGSPIDYGVIVDTKAYSGGYNLPIGQADEMQRYVEENQTRNKHINPNEWWKVYPSSVTEFKFLFVSGHFKGNYKAQLTRLNHITNCNGAVLSVEELLIGGEMIKAGTLTLEEVRRKFNNGEINFRS
ATGGACTACAAAGACCATGACGGTGATTATAAAGATCATGACATCGATTACAAGGATGACGATGACAAGATGGCCCCCAAGAAGAAGAGGAAGGTGGGCATTCACCGCGGGGTACCTATGGTGGACTTGAGGACACTCGGTTATTCGCAACAGCAACAGGAGAAAATCAAGCCTAAGGTCAGGAGCACCGTCGCGCAACACCACGAGGCGCTTGTGGGGCATGGCTTCACTCATGCGCATATTGTCGCGCTTTCACAGCACCCTGCGGCGCTTGGGACGGTGGCTGTCAAATACCAAGATATGATTGCGGCCCTGCCCGAAGCCACGCACGAGGCAATTGTAGGGGTCGGTAAACAGTGGTCGGGAGCGCGAGCACTTGAGGCGCTGCTGACTGTGGCGGGTGAGCTTAGGGGGCCTCCGCTCCAGCTCGACACCGGGCAGCTGCTGAAGATCGCGAAGAGAGGGGGAGTAACAGCGGTAGAGGCAGTGCACGCCTGGCGCAATGCGCTCACCGGGGCCCCCTTGAACCTGACCCCAGACCAGGTAGTCGCAATCGCGTCGCATGACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCAAACGGAGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCAATGGGGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCCACGACGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGACTCACCCCAGACCAGGTAGTCGCAATCGCGTCGAACATTGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCAAACGGAGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCAATGGGGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCAATATTGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGGCTGACCCCAGACCAGGTAGTCGCAATCGCGTCGCATGACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCAAATATCGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCCATGATGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCCACGACGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGCCTGACACCCGAACAGGTGGTCGCCATTGCTAGCAACGGGGGAGGACGGCCAGCCTTGGAGTCCATCGTAGCCCAATTGTCCAGGCCCGATCCCGCGTTGGCTGCGTTAACGAATGACCATCTGGTGGCGTTGGCATGTCTTGGTGGATCCCAGCTGGTGAAGAGCGAGCTGGAGGAGAAGAAGTCCGAGCTGCGGCACAAGCTGAAGTACGTGCCCCACGAGTACATCGAGCTGATCGAGATCGCCAGGAACAGCACCCAGGACCGCATCCTGGAGATGAAGGTGATGGAGTTCTTCATGAAGGTGTACGGCTACAGGGGAAAGCACCTGGGCGGAAGCAGAAAGCCTGACGGCGCCATCTATACAGTGGGCAGCCCCATCGATTACGGCGTGATCGTGGACACAAAGGCCTACAGCGGCGGCTACAATCTGCCTATCGGCCAGGCCGACGAGATGCAGAGATACGTGGAGGAGAACCAGACCCGGAATAAGCACATCAACCCCAACGAGTGGTGGAAGGTGTACCCTAGCAGCGTGACCGAGTTCAAGTTCCTGTTCGTGAGCGGCCACTTCAAGGGCAACTACAAGGCCCAGCTGACCAGGCTGAACCACATCACCAACTGCAATGGCGCCGTGCTGAGCGTGGAGGAGCTGCTGATCGGCGGCGAGATGATCAAAGCCGGCACCCTGACACTGGAGGAGGTGCGGCGCAAGTTCAACAACGGCGAGATCAACTTCAGATCT
MDYKDHDGDYKDHDIDYKDDDDKMAPKKKRKVGIHRGVPMVDLRTLGYSQQQQEKIKPKVRSTVAQHHEALVGHGFTHAHIVALSQHPAALGTVAVKYQDMIAALPEATHEAIVGVGKQWSGARALEALLTVAGELRGPPLQLDTGQLLKIAKRGGVTAVEAVHAWRNALTGAPLNLTPDQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPAQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPDQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPAQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGRPALESIVAQLSRPDPALAALTNDHLVALACLGGRPALDAVKKGLPHAPALIKRTNRRIPERTSHRVAGSQLVKSELEEKKSELRHKLKYVPHEYIELIEIARNSTQDRILEMKVMEFFMKVYGYRGKHLGGSRKPDGAIYTVGSPIDYGVIVDTKAYSGGYNLPIGQADEMQRYVEENQTRNKHINPNEWWKVYPSSVTEFKFLFVSGHFKGNYKAQLTRLNHITNCNGAVLSVEELLIGGEMIKAGTLTLEEVRRKFNNGEINFRS
ATGGACTACAAAGACCATGACGGTGATTATAAAGATCATGACATCGATTACAAGGATGACGATGACAAGATGGCCCCCAAGAAGAAGAGGAAGGTGGGCATTCACCGCGGGGTACCTATGGTGGACTTGAGGACACTCGGTTATTCGCAACAGCAACAGGAGAAAATCAAGCCTAAGGTCAGGAGCACCGTCGCGCAACACCACGAGGCGCTTGTGGGGCATGGCTTCACTCATGCGCATATTGTCGCGCTTTCACAGCACCCTGCGGCGCTTGGGACGGTGGCTGTCAAATACCAAGATATGATTGCGGCCCTGCCCGAAGCCACGCACGAGGCAATTGTAGGGGTCGGTAAACAGTGGTCGGGAGCGCGAGCACTTGAGGCGCTGCTGACTGTGGCGGGTGAGCTTAGGGGGCCTCCGCTCCAGCTCGACACCGGGCAGCTGCTGAAGATCGCGAAGAGAGGGGGAGTAACAGCGGTAGAGGCAGTGCACGCCTGGCGCAATGCGCTCACCGGGGCCCCCTTGAACCTGACCCCAGACCAGGTAGTCGCAATCGCGTCGCATGACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCAAACGGAGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCAATGGGGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCCACGACGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGACTCACCCCAGACCAGGTAGTCGCAATCGCGTCGAACATTGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCAAACGGAGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCAATGGGGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCAATATTGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGGCTGACCCCAGACCAGGTAGTCGCAATCGCGTCGCATGACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCAAATATCGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCCATGATGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCCACGACGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGCCTGACACCCGAACAGGTGGTCGCCATTGCTAGCAACGGGGGAGGACGGCCAGCCTTGGAGTCCATCGTAGCCCAATTGTCCAGGCCCGATCCCGCGTTGGCTGCGTTAACGAATGACCATCTGGTGGCGTTGGCATGTCTTGGTGGACGACCCGCGCTCGATGCAGTCAAAAAGGGTCTGCCTCATGCTCCCGCATTGATCAAAAGAACCAACCGGCGGATTCCCGAGAGAACTTCCCATCGAGTCGCGGGATCCCAGCTGGTGAAGAGCGAGCTGGAGGAGAAGAAGTCCGAGCTGCGGCACAAGCTGAAGTACGTGCCCCACGAGTACATCGAGCTGATCGAGATCGCCAGGAACAGCACCCAGGACCGCATCCTGGAGATGAAGGTGATGGAGTTCTTCATGAAGGTGTACGGCTACAGGGGAAAGCACCTGGGCGGAAGCAGAAAGCCTGACGGCGCCATCTATACAGTGGGCAGCCCCATCGATTACGGCGTGATCGTGGACACAAAGGCCTACAGCGGCGGCTACAATCTGCCTATCGGCCAGGCCGACGAGATGCAGAGATACGTGGAGGAGAACCAGACCCGGAATAAGCACATCAACCCCAACGAGTGGTGGAAGGTGTACCCTAGCAGCGTGACCGAGTTCAAGTTCCTGTTCGTGAGCGGCCACTTCAAGGGCAACTACAAGGCCCAGCTGACCAGGCTGAACCACATCACCAACTGCAATGGCGCCGTGCTGAGCGTGGAGGAGCTGCTGATCGGCGGCGAGATGATCAAAGCCGGCACCCTGACACTGGAGGAGGTGCGGCGCAAGTTCAACAACGGCGAGATCAACTTCAGATCT
MDYKDHDGDYKDHDIDYKDDDDKMAPKKKRKVGIHRGVPMVDLRTLGYSQQQQEKIKPKVRSTVAQHHEALVGHGFTHAHIVALSQHPAALGTVAVKYQDMIAALPEATHEAIVGVGKQWSGARALEALLTVAGELRGPPLQLDTGQLLKIAKRGGVTAVEAVHAWRNALTGAPLNLTPDQVVAIASNGGGKQALETVQRLLPVLCQDHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNGGGKQALETVQRLLPVLCQDHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPAQVVAIANNNGGKQALETVQRLLPVLCQDHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPDQVVAIANNNGGKQALETVQRLLPVLCQAHGLTPAQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGRPALESIVAQLSRPDPALAALTNDHLVALACLGGSQLVKSELEEKKSELRHKLKYVPHEYIELIEIARNSTQDRILEMKVMEFFMKVYGYRGKHLGGSRKPDGAIYTVGSPIDYGVIVDTKAYSGGYNLPIGQADEMQRYVEENQTRNKHINPNEWWKVYPSSVTEFKFLFVSGHFKGNYKAQLTRLNHITNCNGAVLSVEELLIGGEMIKAGTLTLEEVRRKFNNGEINFRS
ATGGACTACAAAGACCATGACGGTGATTATAAAGATCATGACATCGATTACAAGGATGACGATGACAAGATGGCCCCCAAGAAGAAGAGGAAGGTGGGCATTCACCGCGGGGTACCTATGGTGGACTTGAGGACACTCGGTTATTCGCAACAGCAACAGGAGAAAATCAAGCCTAAGGTCAGGAGCACCGTCGCGCAACACCACGAGGCGCTTGTGGGGCATGGCTTCACTCATGCGCATATTGTCGCGCTTTCACAGCACCCTGCGGCGCTTGGGACGGTGGCTGTCAAATACCAAGATATGATTGCGGCCCTGCCCGAAGCCACGCACGAGGCAATTGTAGGGGTCGGTAAACAGTGGTCGGGAGCGCGAGCACTTGAGGCGCTGCTGACTGTGGCGGGTGAGCTTAGGGGGCCTCCGCTCCAGCTCGACACCGGGCAGCTGCTGAAGATCGCGAAGAGAGGGGGAGTAACAGCGGTAGAGGCAGTGCACGCCTGGCGCAATGCGCTCACCGGGGCCCCCTTGAACCTGACCCCAGACCAGGTAGTCGCAATCGCGTCGAATGGCGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCACATGACGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCAACATCGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCAACGGTGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGACTCACCCCAGACCAGGTAGTCGCAATCGCGTCGAATGGCGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCAAATATCGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCCATGATGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCAATATTGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGACCATGGGCTGACCCCAGACCAGGTAGTCGCAATCGCGTCGCATGACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCACATGACGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCAATGGGGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCAACAACAACGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGTTTGACCCCAGACCAGGTAGTCGCAATCGCGTCGCATGACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCAAATATCGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAATAACAATGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCCACGACGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGCCTGACACCCGAACAGGTGGTCGCCATTGCTAGCAACGGGGGAGGACGGCCAGCCTTGGAGTCCATCGTAGCCCAATTGTCCAGGCCCGATCCCGCGTTGGCTGCGTTAACGAATGACCATCTGGTGGCGTTGGCATGTCTTGGTGGATCCCAGCTGGTGAAGAGCGAGCTGGAGGAGAAGAAGTCCGAGCTGCGGCACAAGCTGAAGTACGTGCCCCACGAGTACATCGAGCTGATCGAGATCGCCAGGAACAGCACCCAGGACCGCATCCTGGAGATGAAGGTGATGGAGTTCTTCATGAAGGTGTACGGCTACAGGGGAAAGCACCTGGGCGGAAGCAGAAAGCCTGACGGCGCCATCTATACAGTGGGCAGCCCCATCGATTACGGCGTGATCGTGGACACAAAGGCCTACAGCGGCGGCTACAATCTGCCTATCGGCCAGGCCGACGAGATGCAGAGATACGTGGAGGAGAACCAGACCCGGAATAAGCACATCAACCCCAACGAGTGGTGGAAGGTGTACCCTAGCAGCGTGACCGAGTTCAAGTTCCTGTTCGTGAGCGGCCACTTCAAGGGCAACTACAAGGCCCAGCTGACCAGGCTGAACCACATCACCAACTGCAATGGCGCCGTGCTGAGCGTGGAGGAGCTGCTGATCGGCGGCGAGATGATCAAAGCCGGCACCCTGACACTGGAGGAGGTGCGGCGCAAGTTCAACAACGGCGAGATCAACTTCAGATCT
MDYKDHDGDYKDHDIDYKDDDDKMAPKKKRKVGIHRGVPMVDLRTLGYSQQQQEKIKPKVRSTVAQHHEALVGHGFTHAHIVALSQHPAALGTVAVKYQDMIAALPEATHEAIVGVGKQWSGARALEALLTVAGELRGPPLQLDTGQLLKIAKRGGVTAVEAVHAWRNALTGAPLNLTPDQVVAIASNGGGKQALETVQRLLPVLCQDHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNGGGKQALETVQRLLPVLCQDHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPAQVVAIANNNGGKQALETVQRLLPVLCQDHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPDQVVAIANNNGGKQALETVQRLLPVLCQAHGLTPAQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGRPALESIVAQLSRPDPALAALTNDHLVALACLGGRPALDAVKKGLPHAPALIKRTNRRIPERTSHRVAGSQLVKSELEEKKSELRHKLKYVPHEYIELIEIARNSTQDRILEMKVMEFFMKVYGYRGKHLGGSRKPDGAIYTVGSPIDYGVIVDTKAYSGGYNLPIGQADEMQRYVEENQTRNKHINPNEWWKVYPSSVTEFKFLFVSGHFKGNYKAQLTRLNHITNCNGAVLSVEELLIGGEMIKAGTLTLEEVRRKFNNGEINFRS
ATGGACTACAAAGACCATGACGGTGATTATAAAGATCATGACATCGATTACAAGGATGACGATGACAAGATGGCCCCCAAGAAGAAGAGGAAGGTGGGCATTCACCGCGGGGTACCTATGGTGGACTTGAGGACACTCGGTTATTCGCAACAGCAACAGGAGAAAATCAAGCCTAAGGTCAGGAGCACCGTCGCGCAACACCACGAGGCGCTTGTGGGGCATGGCTTCACTCATGCGCATATTGTCGCGCTTTCACAGCACCCTGCGGCGCTTGGGACGGTGGCTGTCAAATACCAAGATATGATTGCGGCCCTGCCCGAAGCCACGCACGAGGCAATTGTAGGGGTCGGTAAACAGTGGTCGGGAGCGCGAGCACTTGAGGCGCTGCTGACTGTGGCGGGTGAGCTTAGGGGGCCTCCGCTCCAGCTCGACACCGGGCAGCTGCTGAAGATCGCGAAGAGAGGGGGAGTAACAGCGGTAGAGGCAGTGCACGCCTGGCGCAATGCGCTCACCGGGGCCCCCTTGAACCTGACCCCAGACCAGGTAGTCGCAATCGCGTCGAATGGCGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCACATGACGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCAACATCGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCAACGGTGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGACTCACCCCAGACCAGGTAGTCGCAATCGCGTCGAATGGCGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCAAATATCGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCCATGATGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCAATATTGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGGCTGACCCCAGACCAGGTAGTCGCAATCGCGTCGCATGACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCACATGACGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCAATGGGGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCAACAACAACGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGTTTGACCCCAGACCAGGTAGTCGCAATCGCGTCGCATGACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCAAATATCGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAATAACAATGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCCACGACGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGCCTGACACCCGAACAGGTGGTCGCCATTGCTAGCAACGGGGGAGGACGGCCAGCCTTGGAGTCCATCGTAGCCCAATTGTCCAGGCCCGATCCCGCGTTGGCTGCGTTAACGAATGACCATCTGGTGGCGTTGGCATGTCTTGGTGGACGACCCGCGCTCGATGCAGTCAAAAAGGGTCTGCCTCATGCTCCCGCATTGATCAAAAGAACCAACCGGCGGATTCCCGAGAGAACTTCCCATCGAGTCGCGGGATCCCAGCTGGTGAAGAGCGAGCTGGAGGAGAAGAAGTCCGAGCTGCGGCACAAGCTGAAGTACGTGCCCCACGAGTACATCGAGCTGATCGAGATCGCCAGGAACAGCACCCAGGACCGCATCCTGGAGATGAAGGTGATGGAGTTCTTCATGAAGGTGTACGGCTACAGGGGAAAGCACCTGGGCGGAAGCAGAAAGCCTGACGGCGCCATCTATACAGTGGGCAGCCCCATCGATTACGGCGTGATCGTGGACACAAAGGCCTACAGCGGCGGCTACAATCTGCCTATCGGCCAGGCCGACGAGATGCAGAGATACGTGGAGGAGAACCAGACCCGGAATAAGCACATCAACCCCAACGAGTGGTGGAAGGTGTACCCTAGCAGCGTGACCGAGTTCAAGTTCCTGTTCGTGAGCGGCCACTTCAAGGGCAACTACAAGGCCCAGCTGACCAGGCTGAACCACATCACCAACTGCAATGGCGCCGTGCTGAGCGTGGAGGAGCTGCTGATCGGCGGCGAGATGATCAAAGCCGGCACCCTGACACTGGAGGAGGTGCGGCGCAAGTTCAACAACGGCGAGATCAACTTCAGATCT
MDYKDHDGDYKDHDIDYKDDDDKMAPKKKRKVGIHRGVPMVDLRTLGYSQQQQEKIKPKVRSTVAQHHEALVGHGFTHAHIVALSQHPAALGTVAVKYQDMIAALPEATHEAIVGVGKQWSGARALEALLTVAGELRGPPLQLDTGQLLKIAKRGGVTAVEAVHAWRNALTGAPLNLTPDQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPAQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPDQVVAIANNNGGKQALETVQRLLPVLCQAHGLTPAQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPDQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPDQVVAIANNNGGKQALETVQRLLPVLCQAHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNGGGKQALETVQRLLPVLCQDHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGRPALESIVAQLSRPDPALAALTNDHLVALACLGGSQLVKSELEEKKSELRHKLKYVPHEYIELIEIARNSTQDRILEMKVMEFFMKVYGYRGKHLGGSRKPDGAIYTVGSPIDYGVIVDTKAYSGGYNLPIGQADEMQRYVEENQTRNKHINPNEWWKVYPSSVTEFKFLFVSGHFKGNYKAQLTRLNHITNCNGAVLSVEELLIGGEMIKAGTLTLEEVRRKFNNGEINFRS
ATGGACTACAAAGACCATGACGGTGATTATAAAGATCATGACATCGATTACAAGGATGACGATGACAAGATGGCCCCCAAGAAGAAGAGGAAGGTGGGCATTCACCGCGGGGTACCTATGGTGGACTTGAGGACACTCGGTTATTCGCAACAGCAACAGGAGAAAATCAAGCCTAAGGTCAGGAGCACCGTCGCGCAACACCACGAGGCGCTTGTGGGGCATGGCTTCACTCATGCGCATATTGTCGCGCTTTCACAGCACCCTGCGGCGCTTGGGACGGTGGCTGTCAAATACCAAGATATGATTGCGGCCCTGCCCGAAGCCACGCACGAGGCAATTGTAGGGGTCGGTAAACAGTGGTCGGGAGCGCGAGCACTTGAGGCGCTGCTGACTGTGGCGGGTGAGCTTAGGGGGCCTCCGCTCCAGCTCGACACCGGGCAGCTGCTGAAGATCGCGAAGAGAGGGGGAGTAACAGCGGTAGAGGCAGTGCACGCCTGGCGCAATGCGCTCACCGGGGCCCCCTTGAACCTGACCCCAGACCAGGTAGTCGCAATCGCGTCGAACATTGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCACATGACGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCAACATCGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCCACGACGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGACTCACCCCAGACCAGGTAGTCGCAATCGCGTCGCATGACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCAAACGGAGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAATAACAATGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCCACGACGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGGCTGACCCCAGACCAGGTAGTCGCAATCGCGTCGAACATTGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGATCAAGTCGTGGCCATTGCAAATAATAACGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCCATGATGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCAACGGTGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGTTTGACCCCAGACCAGGTAGTCGCAATCGCGTCGCATGACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTACTGTGCCAGGATCATGGCCTGACACCCGAACAGGTGGTCGCCATTGCTAGCAACGGGGGAGGACGGCCAGCCTTGGAGTCCATCGTAGCCCAATTGTCCAGGCCCGATCCCGCGTTGGCTGCGTTAACGAATGACCATCTGGTGGCGTTGGCATGTCTTGGTGGATCCCAGCTGGTGAAGAGCGAGCTGGAGGAGAAGAAGTCCGAGCTGCGGCACAAGCTGAAGTACGTGCCCCACGAGTACATCGAGCTGATCGAGATCGCCAGGAACAGCACCCAGGACCGCATCCTGGAGATGAAGGTGATGGAGTTCTTCATGAAGGTGTACGGCTACAGGGGAAAGCACCTGGGCGGAAGCAGAAAGCCTGACGGCGCCATCTATACAGTGGGCAGCCCCATCGATTACGGCGTGATCGTGGACACAAAGGCCTACAGCGGCGGCTACAATCTGCCTATCGGCCAGGCCGACGAGATGCAGAGATACGTGGAGGAGAACCAGACCCGGAATAAGCACATCAACCCCAACGAGTGGTGGAAGGTGTACCCTAGCAGCGTGACCGAGTTCAAGTTCCTGTTCGTGAGCGGCCACTTCAAGGGCAACTACAAGGCCCAGCTGACCAGGCTGAACCACATCACCAACTGCAATGGCGCCGTGCTGAGCGTGGAGGAGCTGCTGATCGGCGGCGAGATGATCAAAGCCGGCACCCTGACACTGGAGGAGGTGCGGCGCAAGTTCAACAACGGCGAGATCAACTTCAGATCT
MDYKDHDGDYKDHDIDYKDDDDKMAPKKKRKVGIHRGVPMVDLRTLGYSQQQQEKIKPKVRSTVAQHHEALVGHGFTHAHIVALSQHPAALGTVAVKYQDMIAALPEATHEAIVGVGKQWSGARALEALLTVAGELRGPPLQLDTGQLLKIAKRGGVTAVEAVHAWRNALTGAPLNLTPDQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPAQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPDQVVAIANNNGGKQALETVQRLLPVLCQAHGLTPAQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPDQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPDQVVAIANNNGGKQALETVQRLLPVLCQAHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNGGGKQALETVQRLLPVLCQDHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGRPALESIVAQLSRPDPALAALTNDHLVALACLGGRPALDAVKKGLPHAPALIKRTNRRIPERTSHRVAGSQLVKSELEEKKSELRHKLKYVPHEYIELIEIARNSTQDRILEMKVMEFFMKVYGYRGKHLGGSRKPDGAIYTVGSPIDYGVIVDTKAYSGGYNLPIGQADEMQRYVEENQTRNKHINPNEWWKVYPSSVTEFKFLFVSGHFKGNYKAQLTRLNHITNCNGAVLSVEELLIGGEMIKAGTLTLEEVRRKFNNGEINFRS
ATGGACTACAAAGACCATGACGGTGATTATAAAGATCATGACATCGATTACAAGGATGACGATGACAAGATGGCCCCCAAGAAGAAGAGGAAGGTGGGCATTCACCGCGGGGTACCTATGGTGGACTTGAGGACACTCGGTTATTCGCAACAGCAACAGGAGAAAATCAAGCCTAAGGTCAGGAGCACCGTCGCGCAACACCACGAGGCGCTTGTGGGGCATGGCTTCACTCATGCGCATATTGTCGCGCTTTCACAGCACCCTGCGGCGCTTGGGACGGTGGCTGTCAAATACCAAGATATGATTGCGGCCCTGCCCGAAGCCACGCACGAGGCAATTGTAGGGGTCGGTAAACAGTGGTCGGGAGCGCGAGCACTTGAGGCGCTGCTGACTGTGGCGGGTGAGCTTAGGGGGCCTCCGCTCCAGCTCGACACCGGGCAGCTGCTGAAGATCGCGAAGAGAGGGGGAGTAACAGCGGTAGAGGCAGTGCACGCCTGGCGCAATGCGCTCACCGGGGCCCCCTTGAACCTGACCCCAGACCAGGTAGTCGCAATCGCGTCGAACATTGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCACATGACGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCAACATCGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCCACGACGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGACTCACCCCAGACCAGGTAGTCGCAATCGCGTCGCATGACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCAAACGGAGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAATAACAATGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCCACGACGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGGCTGACCCCAGACCAGGTAGTCGCAATCGCGTCGAACATTGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGATCAAGTCGTGGCCATTGCAAATAATAACGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCCATGATGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCAACGGTGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGTTTGACCCCAGACCAGGTAGTCGCAATCGCGTCGCATGACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTACTGTGCCAGGATCATGGCCTGACACCCGAACAGGTGGTCGCCATTGCTAGCAACGGGGGAGGACGGCCAGCCTTGGAGTCCATCGTAGCCCAATTGTCCAGGCCCGATCCCGCGTTGGCTGCGTTAACGAATGACCATCTGGTGGCGTTGGCATGTCTTGGTGGACGACCCGCGCTCGATGCAGTCAAAAAGGGTCTGCCTCATGCTCCCGCATTGATCAAAAGAACCAACCGGCGGATTCCCGAGAGAACTTCCCATCGAGTCGCGGGATCCCAGCTGGTGAAGAGCGAGCTGGAGGAGAAGAAGTCCGAGCTGCGGCACAAGCTGAAGTACGTGCCCCACGAGTACATCGAGCTGATCGAGATCGCCAGGAACAGCACCCAGGACCGCATCCTGGAGATGAAGGTGATGGAGTTCTTCATGAAGGTGTACGGCTACAGGGGAAAGCACCTGGGCGGAAGCAGAAAGCCTGACGGCGCCATCTATACAGTGGGCAGCCCCATCGATTACGGCGTGATCGTGGACACAAAGGCCTACAGCGGCGGCTACAATCTGCCTATCGGCCAGGCCGACGAGATGCAGAGATACGTGGAGGAGAACCAGACCCGGAATAAGCACATCAACCCCAACGAGTGGTGGAAGGTGTACCCTAGCAGCGTGACCGAGTTCAAGTTCCTGTTCGTGAGCGGCCACTTCAAGGGCAACTACAAGGCCCAGCTGACCAGGCTGAACCACATCACCAACTGCAATGGCGCCGTGCTGAGCGTGGAGGAGCTGCTGATCGGCGGCGAGATGATCAAAGCCGGCACCCTGACACTGGAGGAGGTGCGGCGCAAGTTCAACAACGGCGAGATCAACTTCAGATCT
MDYKDHDGDYKDHDIDYKDDDDKMAPKKKRKVGIHRGVPMVDLRTLGYSQQQQEKIKPKVRSTVAQHHEALVGHGFTHAHIVALSQHPAALGTVAVKYQDMIAALPEATHEAIVGVGKQWSGARALEALLTVAGELRGPPLQLDTGQLLKIAKRGGVTAVEAVHAWRNALTGAPLNLTPDQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPAQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPDQVVAIANNNGGKQALETVQRLLPVLCQAHGLTPAQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPDQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPDQVVAIANNNGGKQALETVQRLLPVLCQAHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNGGGKQALETVQRLLPVLCQDHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGRPALESIVAQLSRPDPALAALTNDHLVALACLGGSQLVKSELEEKKSELRHKLKYVPHEYIELIEIARNSTQDRILEMKVMEFFMKVYGYRGKHLGGSRKPDGAIYTVGSPIDYGVIVDTKAYSGGYNLPIGQADEMQRYVEENQTRNKHINPNEWWKVYPSSVTEFKFLFVSGHFKGNYKAQLTRLNHITNCNGAVLSVEELLIGGEMIKAGTLTLEEVRRKFNNGEINFRS
ATGGACTACAAAGACCATGACGGTGATTATAAAGATCATGACATCGATTACAAGGATGACGATGACAAGATGGCCCCCAAGAAGAAGAGGAAGGTGGGCATTCACCGCGGGGTACCTATGGTGGACTTGAGGACACTCGGTTATTCGCAACAGCAACAGGAGAAAATCAAGCCTAAGGTCAGGAGCACCGTCGCGCAACACCACGAGGCGCTTGTGGGGCATGGCTTCACTCATGCGCATATTGTCGCGCTTTCACAGCACCCTGCGGCGCTTGGGACGGTGGCTGTCAAATACCAAGATATGATTGCGGCCCTGCCCGAAGCCACGCACGAGGCAATTGTAGGGGTCGGTAAACAGTGGTCGGGAGCGCGAGCACTTGAGGCGCTGCTGACTGTGGCGGGTGAGCTTAGGGGGCCTCCGCTCCAGCTCGACACCGGGCAGCTGCTGAAGATCGCGAAGAGAGGGGGAGTAACAGCGGTAGAGGCAGTGCACGCCTGGCGCAATGCGCTCACCGGGGCCCCCTTGAACCTGACCCCAGACCAGGTAGTCGCAATCGCGTCGAACATTGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCACATGACGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCAACATCGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCCACGACGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGACTCACCCCAGACCAGGTAGTCGCAATCGCGTCGCATGACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCAAACGGAGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAATAACAATGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCCACGACGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGGCTGACCCCAGACCAGGTAGTCGCAATCGCGTCGAACATTGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGATCAAGTCGTGGCCATTGCAAATAATAACGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCCATGATGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCAACGGTGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGTTTGACCCCAGACCAGGTAGTCGCAATCGCGTCGCATGACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCAAACGGAGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCCATGATGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCAATATTGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGCCTGACACCCGAACAGGTGGTCGCCATTGCTAGCAACGGGGGAGGACGGCCAGCCTTGGAGTCCATCGTAGCCCAATTGTCCAGGCCCGATCCCGCGTTGGCTGCGTTAACGAATGACCATCTGGTGGCGTTGGCATGTCTTGGTGGATCCCAGCTGGTGAAGAGCGAGCTGGAGGAGAAGAAGTCCGAGCTGCGGCACAAGCTGAAGTACGTGCCCCACGAGTACATCGAGCTGATCGAGATCGCCAGGAACAGCACCCAGGACCGCATCCTGGAGATGAAGGTGATGGAGTTCTTCATGAAGGTGTACGGCTACAGGGGAAAGCACCTGGGCGGAAGCAGAAAGCCTGACGGCGCCATCTATACAGTGGGCAGCCCCATCGATTACGGCGTGATCGTGGACACAAAGGCCTACAGCGGCGGCTACAATCTGCCTATCGGCCAGGCCGACGAGATGCAGAGATACGTGGAGGAGAACCAGACCCGGAATAAGCACATCAACCCCAACGAGTGGTGGAAGGTGTACCCTAGCAGCGTGACCGAGTTCAAGTTCCTGTTCGTGAGCGGCCACTTCAAGGGCAACTACAAGGCCCAGCTGACCAGGCTGAACCACATCACCAACTGCAATGGCGCCGTGCTGAGCGTGGAGGAGCTGCTGATCGGCGGCGAGATGATCAAAGCCGGCACCCTGACACTGGAGGAGGTGCGGCGCAAGTTCAACAACGGCGAGATCAACTTCAGATCT
MDYKDHDGDYKDHDIDYKDDDDKMAPKKKRKVGIHRGVPMVDLRTLGYSQQQQEKIKPKVRSTVAQHHEALVGHGFTHAHIVALSQHPAALGTVAVKYQDMIAALPEATHEAIVGVGKQWSGARALEALLTVAGELRGPPLQLDTGQLLKIAKRGGVTAVEAVHAWRNALTGAPLNLTPDQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPAQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPDQVVAIANNNGGKQALETVQRLLPVLCQAHGLTPAQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPDQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPDQVVAIANNNGGKQALETVQRLLPVLCQAHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNGGGKQALETVQRLLPVLCQDHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGRPALESIVAQLSRPDPALAALTNDHLVALACLGGRPALDAVKKGLPHAPALIKRTNRRIPERTSHRVAGSQLVKSELEEKKSELRHKLKYVPHEYIELIEIARNSTQDRILEMKVMEFFMKVYGYRGKHLGGSRKPDGAIYTVGSPIDYGVIVDTKAYSGGYNLPIGQADEMQRYVEENQTRNKHINPNEWWKVYPSSVTEFKFLFVSGHFKGNYKAQLTRLNHITNCNGAVLSVEELLIGGEMIKAGTLTLEEVRRKFNNGEINFRS
ATGGACTACAAAGACCATGACGGTGATTATAAAGATCATGACATCGATTACAAGGATGACGATGACAAGATGGCCCCCAAGAAGAAGAGGAAGGTGGGCATTCACCGCGGGGTACCTATGGTGGACTTGAGGACACTCGGTTATTCGCAACAGCAACAGGAGAAAATCAAGCCTAAGGTCAGGAGCACCGTCGCGCAACACCACGAGGCGCTTGTGGGGCATGGCTTCACTCATGCGCATATTGTCGCGCTTTCACAGCACCCTGCGGCGCTTGGGACGGTGGCTGTCAAATACCAAGATATGATTGCGGCCCTGCCCGAAGCCACGCACGAGGCAATTGTAGGGGTCGGTAAACAGTGGTCGGGAGCGCGAGCACTTGAGGCGCTGCTGACTGTGGCGGGTGAGCTTAGGGGGCCTCCGCTCCAGCTCGACACCGGGCAGCTGCTGAAGATCGCGAAGAGAGGGGGAGTAACAGCGGTAGAGGCAGTGCACGCCTGGCGCAATGCGCTCACCGGGGCCCCCTTGAACCTGACCCCAGACCAGGTAGTCGCAATCGCGTCGAACATTGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCACATGACGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCAACATCGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCCACGACGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGACTCACCCCAGACCAGGTAGTCGCAATCGCGTCGCATGACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCAAACGGAGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAATAACAATGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCCACGACGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGGCTGACCCCAGACCAGGTAGTCGCAATCGCGTCGAACATTGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGATCAAGTCGTGGCCATTGCAAATAATAACGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCCATGATGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCAACGGTGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGTTTGACCCCAGACCAGGTAGTCGCAATCGCGTCGCATGACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCAAACGGAGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCCATGATGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCAATATTGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGCCTGACACCCGAACAGGTGGTCGCCATTGCTAGCAACGGGGGAGGACGGCCAGCCTTGGAGTCCATCGTAGCCCAATTGTCCAGGCCCGATCCCGCGTTGGCTGCGTTAACGAATGACCATCTGGTGGCGTTGGCATGTCTTGGTGGACGACCCGCGCTCGATGCAGTCAAAAAGGGTCTGCCTCATGCTCCCGCATTGATCAAAAGAACCAACCGGCGGATTCCCGAGAGAACTTCCCATCGAGTCGCGGGATCCCAGCTGGTGAAGAGCGAGCTGGAGGAGAAGAAGTCCGAGCTGCGGCACAAGCTGAAGTACGTGCCCCACGAGTACATCGAGCTGATCGAGATCGCCAGGAACAGCACCCAGGACCGCATCCTGGAGATGAAGGTGATGGAGTTCTTCATGAAGGTGTACGGCTACAGGGGAAAGCACCTGGGCGGAAGCAGAAAGCCTGACGGCGCCATCTATACAGTGGGCAGCCCCATCGATTACGGCGTGATCGTGGACACAAAGGCCTACAGCGGCGGCTACAATCTGCCTATCGGCCAGGCCGACGAGATGCAGAGATACGTGGAGGAGAACCAGACCCGGAATAAGCACATCAACCCCAACGAGTGGTGGAAGGTGTACCCTAGCAGCGTGACCGAGTTCAAGTTCCTGTTCGTGAGCGGCCACTTCAAGGGCAACTACAAGGCCCAGCTGACCAGGCTGAACCACATCACCAACTGCAATGGCGCCGTGCTGAGCGTGGAGGAGCTGCTGATCGGCGGCGAGATGATCAAAGCCGGCACCCTGACACTGGAGGAGGTGCGGCGCAAGTTCAACAACGGCGAGATCAACTTCAGATCT
MDYKDHDGDYKDHDIDYKDDDDKMAPKKKRKVGIHRGVPMVDLRTLGYSQQQQEKIKPKVRSTVAQHHEALVGHGFTHAHIVALSQHPAALGTVAVKYQDMIAALPEATHEAIVGVGKQWSGARALEALLTVAGELRGPPLQLDTGQLLKIAKRGGVTAVEAVHAWRNALTGAPLNLTPDQVVAIANNNGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPDQVVAIANNNGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPDQVVAIANNNGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNGGGKQALETVQRLLPVLCQDHGLTPDQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGRPALESIVAQLSRPDPALAALTNDHLVALACLGGSQLVKSELEEKKSELRHKLKYVPHEYIELIEIARNSTQDRILEMKVMEFFMKVYGYRGKHLGGSRKPDGAIYTVGSPIDYGVIVDTKAYSGGYNLPIGQADEMQRYVEENQTRNKHINPNEWWKVYPSSVTEFKFLFVSGHFKGNYKAQLTRLNHITNCNGAVLSVEELLIGGEMIKAGTLTLEEVRRKFNNGEINFRS
ATGGACTACAAAGACCATGACGGTGATTATAAAGATCATGACATCGATTACAAGGATGACGATGACAAGATGGCCCCCAAGAAGAAGAGGAAGGTGGGCATTCACCGCGGGGTACCTATGGTGGACTTGAGGACACTCGGTTATTCGCAACAGCAACAGGAGAAAATCAAGCCTAAGGTCAGGAGCACCGTCGCGCAACACCACGAGGCGCTTGTGGGGCATGGCTTCACTCATGCGCATATTGTCGCGCTTTCACAGCACCCTGCGGCGCTTGGGACGGTGGCTGTCAAATACCAAGATATGATTGCGGCCCTGCCCGAAGCCACGCACGAGGCAATTGTAGGGGTCGGTAAACAGTGGTCGGGAGCGCGAGCACTTGAGGCGCTGCTGACTGTGGCGGGTGAGCTTAGGGGGCCTCCGCTCCAGCTCGACACCGGGCAGCTGCTGAAGATCGCGAAGAGAGGGGGAGTAACAGCGGTAGAGGCAGTGCACGCCTGGCGCAATGCGCTCACCGGGGCCCCCTTGAACCTGACCCCAGACCAGGTAGTCGCAATCGCCAACAATAACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCAAATATCGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCAATGGGGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCAATATTGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGACTCACCCCAGACCAGGTAGTCGCAATCGCGTCGCATGACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCAAACGGAGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAATAACAATGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCAATATTGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGGCTGACCCCAGACCAGGTAGTCGCAATCGCGTCGCATGACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCAAACGGAGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAATAACAATGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCAACGGTGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGTTTGACCCCAGACCAGGTAGTCGCAATCGCGTCGAACATTGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTACTGTGCCAGGATCATGGCCTGACACCCGAACAGGTGGTCGCCATTGCTAGCAACGGGGGAGGACGGCCAGCCTTGGAGTCCATCGTAGCCCAATTGTCCAGGCCCGATCCCGCGTTGGCTGCGTTAACGAATGACCATCTGGTGGCGTTGGCATGTCTTGGTGGATCCCAGCTGGTGAAGAGCGAGCTGGAGGAGAAGAAGTCCGAGCTGCGGCACAAGCTGAAGTACGTGCCCCACGAGTACATCGAGCTGATCGAGATCGCCAGGAACAGCACCCAGGACCGCATCCTGGAGATGAAGGTGATGGAGTTCTTCATGAAGGTGTACGGCTACAGGGGAAAGCACCTGGGCGGAAGCAGAAAGCCTGACGGCGCCATCTATACAGTGGGCAGCCCCATCGATTACGGCGTGATCGTGGACACAAAGGCCTACAGCGGCGGCTACAATCTGCCTATCGGCCAGGCCGACGAGATGCAGAGATACGTGGAGGAGAACCAGACCCGGAATAAGCACATCAACCCCAACGAGTGGTGGAAGGTGTACCCTAGCAGCGTGACCGAGTTCAAGTTCCTGTTCGTGAGCGGCCACTTCAAGGGCAACTACAAGGCCCAGCTGACCAGGCTGAACCACATCACCAACTGCAATGGCGCCGTGCTGAGCGTGGAGGAGCTGCTGATCGGCGGCGAGATGATCAAAGCCGGCACCCTGACACTGGAGGAGGTGCGGCGCAAGTTCAACAACGGCGAGATCAACTTCAGATCT
MDYKDHDGDYKDHDIDYKDDDDKMAPKKKRKVGIHRGVPMVDLRTLGYSQQQQEKIKPKVRSTVAQHHEALVGHGFTHAHIVALSQHPAALGTVAVKYQDMIAALPEATHEAIVGVGKQWSGARALEALLTVAGELRGPPLQLDTGQLLKIAKRGGVTAVEAVHAWRNALTGAPLNLTPDQVVAIANNNGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPDQVVAIANNNGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPDQVVAIANNNGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNGGGKQALETVQRLLPVLCQDHGLTPDQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGRPALESIVAQLSRPDPALAALTNDHLVALACLGGRPALDAVKKGLPHAPALIKRTNRRIPERTSHRVAGSQLVKSELEEKKSELRHKLKYVPHEYIELIEIARNSTQDRILEMKVMEFFMKVYGYRGKHLGGSRKPDGAIYTVGSPIDYGVIVDTKAYSGGYNLPIGQADEMQRYVEENQTRNKHINPNEWWKVYPSSVTEFKFLFVSGHFKGNYKAQLTRLNHITNCNGAVLSVEELLIGGEMIKAGTLTLEEVRRKFNNGEINFRS
ATGGACTACAAAGACCATGACGGTGATTATAAAGATCATGACATCGATTACAAGGATGACGATGACAAGATGGCCCCCAAGAAGAAGAGGAAGGTGGGCATTCACCGCGGGGTACCTATGGTGGACTTGAGGACACTCGGTTATTCGCAACAGCAACAGGAGAAAATCAAGCCTAAGGTCAGGAGCACCGTCGCGCAACACCACGAGGCGCTTGTGGGGCATGGCTTCACTCATGCGCATATTGTCGCGCTTTCACAGCACCCTGCGGCGCTTGGGACGGTGGCTGTCAAATACCAAGATATGATTGCGGCCCTGCCCGAAGCCACGCACGAGGCAATTGTAGGGGTCGGTAAACAGTGGTCGGGAGCGCGAGCACTTGAGGCGCTGCTGACTGTGGCGGGTGAGCTTAGGGGGCCTCCGCTCCAGCTCGACACCGGGCAGCTGCTGAAGATCGCGAAGAGAGGGGGAGTAACAGCGGTAGAGGCAGTGCACGCCTGGCGCAATGCGCTCACCGGGGCCCCCTTGAACCTGACCCCAGACCAGGTAGTCGCAATCGCCAACAATAACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCAAATATCGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCAATGGGGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCAATATTGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGACTCACCCCAGACCAGGTAGTCGCAATCGCGTCGCATGACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCAAACGGAGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAATAACAATGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCAATATTGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGGCTGACCCCAGACCAGGTAGTCGCAATCGCGTCGCATGACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCAAACGGAGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAATAACAATGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCAACGGTGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGTTTGACCCCAGACCAGGTAGTCGCAATCGCGTCGAACATTGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTACTGTGCCAGGATCATGGCCTGACACCCGAACAGGTGGTCGCCATTGCTAGCAACGGGGGAGGACGGCCAGCCTTGGAGTCCATCGTAGCCCAATTGTCCAGGCCCGATCCCGCGTTGGCTGCGTTAACGAATGACCATCTGGTGGCGTTGGCATGTCTTGGTGGACGACCCGCGCTCGATGCAGTCAAAAAGGGTCTGCCTCATGCTCCCGCATTGATCAAAAGAACCAACCGGCGGATTCCCGAGAGAACTTCCCATCGAGTCGCGGGATCCCAGCTGGTGAAGAGCGAGCTGGAGGAGAAGAAGTCCGAGCTGCGGCACAAGCTGAAGTACGTGCCCCACGAGTACATCGAGCTGATCGAGATCGCCAGGAACAGCACCCAGGACCGCATCCTGGAGATGAAGGTGATGGAGTTCTTCATGAAGGTGTACGGCTACAGGGGAAAGCACCTGGGCGGAAGCAGAAAGCCTGACGGCGCCATCTATACAGTGGGCAGCCCCATCGATTACGGCGTGATCGTGGACACAAAGGCCTACAGCGGCGGCTACAATCTGCCTATCGGCCAGGCCGACGAGATGCAGAGATACGTGGAGGAGAACCAGACCCGGAATAAGCACATCAACCCCAACGAGTGGTGGAAGGTGTACCCTAGCAGCGTGACCGAGTTCAAGTTCCTGTTCGTGAGCGGCCACTTCAAGGGCAACTACAAGGCCCAGCTGACCAGGCTGAACCACATCACCAACTGCAATGGCGCCGTGCTGAGCGTGGAGGAGCTGCTGATCGGCGGCGAGATGATCAAAGCCGGCACCCTGACACTGGAGGAGGTGCGGCGCAAGTTCAACAACGGCGAGATCAACTTCAGATCT
MDYKDHDGDYKDHDIDYKDDDDKMAPKKKRKVGIHRGVPMVDLRTLGYSQQQQEKIKPKVRSTVAQHHEALVGHGFTHAHIVALSQHPAALGTVAVKYQDMIAALPEATHEAIVGVGKQWSGARALEALLTVAGELRGPPLQLDTGQLLKIAKRGGVTAVEAVHAWRNALTGAPLNLTPDQVVAIASNGGGKQALETVQRLLPVLCQDHGLTPDQVVAIANNNGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNGGGKQALETVQRLLPVLCQDHGLTPDQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPAQVVAIANNNGGKQALETVQRLLPVLCQDHGLTPDQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNGGGKQTLETVQRLLPVLCQAHGLTPAQVVAIANNNGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGRPALESIVAQLSRPDPALAALTNDHLVALACLGGSQLVKSELEEKKSELRHKLKYVPHEYIELIEIARNSTQDRILEMKVMEFFMKVYGYRGKHLGGSRKPDGAIYTVGSPIDYGVIVDTKAYSGGYNLPIGQADEMQRYVEENQTRNKHINPNEWWKVYPSSVTEFKFLFVSGHFKGNYKAQLTRLNHITNCNGAVLSVEELLIGGEMIKAGTLTLEEVRRKFNNGEINFRS
ATGGACTACAAAGACCATGACGGTGATTATAAAGATCATGACATCGATTACAAGGATGACGATGACAAGATGGCCCCCAAGAAGAAGAGGAAGGTGGGCATTCACCGCGGGGTACCTATGGTGGACTTGAGGACACTCGGTTATTCGCAACAGCAACAGGAGAAAATCAAGCCTAAGGTCAGGAGCACCGTCGCGCAACACCACGAGGCGCTTGTGGGGCATGGCTTCACTCATGCGCATATTGTCGCGCTTTCACAGCACCCTGCGGCGCTTGGGACGGTGGCTGTCAAATACCAAGATATGATTGCGGCCCTGCCCGAAGCCACGCACGAGGCAATTGTAGGGGTCGGTAAACAGTGGTCGGGAGCGCGAGCACTTGAGGCGCTGCTGACTGTGGCGGGTGAGCTTAGGGGGCCTCCGCTCCAGCTCGACACCGGGCAGCTGCTGAAGATCGCGAAGAGAGGGGGAGTAACAGCGGTAGAGGCAGTGCACGCCTGGCGCAATGCGCTCACCGGGGCCCCCTTGAACCTGACCCCAGACCAGGTAGTCGCAATCGCGTCGAATGGCGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGATCAAGTCGTGGCCATTGCAAATAATAACGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCAACATCGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCAACGGTGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGACTCACCCCAGACCAGGTAGTCGCAATCGCGTCGAACATTGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCACATGACGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCAATGGGGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCAACAACAACGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGGCTGACCCCAGACCAGGTAGTCGCAATCGCGTCGAACATTGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCACATGACGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCAATGGGGGAGGGAAACAAACATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCAACAACAACGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGCCTGACACCCGAACAGGTGGTCGCCATTGCTAGCAACGGGGGAGGACGGCCAGCCTTGGAGTCCATCGTAGCCCAATTGTCCAGGCCCGATCCCGCGTTGGCTGCGTTAACGAATGACCATCTGGTGGCGTTGGCATGTCTTGGTGGATCCCAGCTGGTGAAGAGCGAGCTGGAGGAGAAGAAGTCCGAGCTGCGGCACAAGCTGAAGTACGTGCCCCACGAGTACATCGAGCTGATCGAGATCGCCAGGAACAGCACCCAGGACCGCATCCTGGAGATGAAGGTGATGGAGTTCTTCATGAAGGTGTACGGCTACAGGGGAAAGCACCTGGGCGGAAGCAGAAAGCCTGACGGCGCCATCTATACAGTGGGCAGCCCCATCGATTACGGCGTGATCGTGGACACAAAGGCCTACAGCGGCGGCTACAATCTGCCTATCGGCCAGGCCGACGAGATGCAGAGATACGTGGAGGAGAACCAGACCCGGAATAAGCACATCAACCCCAACGAGTGGTGGAAGGTGTACCCTAGCAGCGTGACCGAGTTCAAGTTCCTGTTCGTGAGCGGCCACTTCAAGGGCAACTACAAGGCCCAGCTGACCAGGCTGAACCACATCACCAACTGCAATGGCGCCGTGCTGAGCGTGGAGGAGCTGCTGATCGGCGGCGAGATGATCAAAGCCGGCACCCTGACACTGGAGGAGGTGCGGCGCAAGTTCAACAACGGCGAGATCAACTTCAGATCT
MDYKDHDGDYKDHDIDYKDDDDKMAPKKKRKVGIHRGVPMVDLRTLGYSQQQQEKIKPKVRSTVAQHHEALVGHGFTHAHIVALSQHPAALGTVAVKYQDMIAALPEATHEAIVGVGKQWSGARALEALLTVAGELRGPPLQLDTGQLLKIAKRGGVTAVEAVHAWRNALTGAPLNLTPDQVVAIASNGGGKQALETVQRLLPVLCQDHGLTPDQVVAIANNNGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNGGGKQALETVQRLLPVLCQDHGLTPDQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPAQVVAIANNNGGKQALETVQRLLPVLCQDHGLTPDQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPAQVVAIANNNGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGRPALESIVAQLSRPDPALAALTNDHLVALACLGGRPALDAVKKGLPHAPALIKRTNRRIPERTSHRVAGSQLVKSELEEKKSELRHKLKYVPHEYIELIEIARNSTQDRILEMKVMEFFMKVYGYRGKHLGGSRKPDGAIYTVGSPIDYGVIVDTKAYSGGYNLPIGQADEMQRYVEENQTRNKHINPNEWWKVYPSSVTEFKFLFVSGHFKGNYKAQLTRLNHITNCNGAVLSVEELLIGGEMIKAGTLTLEEVRRKFNNGEINFRS
ATGGACTACAAAGACCATGACGGTGATTATAAAGATCATGACATCGATTACAAGGATGACGATGACAAGATGGCCCCCAAGAAGAAGAGGAAGGTGGGCATTCACCGCGGGGTACCTATGGTGGACTTGAGGACACTCGGTTATTCGCAACAGCAACAGGAGAAAATCAAGCCTAAGGTCAGGAGCACCGTCGCGCAACACCACGAGGCGCTTGTGGGGCATGGCTTCACTCATGCGCATATTGTCGCGCTTTCACAGCACCCTGCGGCGCTTGGGACGGTGGCTGTCAAATACCAAGATATGATTGCGGCCCTGCCCGAAGCCACGCACGAGGCAATTGTAGGGGTCGGTAAACAGTGGTCGGGAGCGCGAGCACTTGAGGCGCTGCTGACTGTGGCGGGTGAGCTTAGGGGGCCTCCGCTCCAGCTCGACACCGGGCAGCTGCTGAAGATCGCGAAGAGAGGGGGAGTAACAGCGGTAGAGGCAGTGCACGCCTGGCGCAATGCGCTCACCGGGGCCCCCTTGAACCTGACCCCAGACCAGGTAGTCGCAATCGCGTCGAATGGCGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGATCAAGTCGTGGCCATTGCAAATAATAACGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCAACATCGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCAACGGTGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGACTCACCCCAGACCAGGTAGTCGCAATCGCGTCGAACATTGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCACATGACGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCAATGGGGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCAACAACAACGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGGCTGACCCCAGACCAGGTAGTCGCAATCGCGTCGAACATTGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCACATGACGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCAATGGGGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCAACAACAACGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGCCTGACACCCGAACAGGTGGTCGCCATTGCTAGCAACGGGGGAGGACGGCCAGCCTTGGAGTCCATCGTAGCCCAATTGTCCAGGCCCGATCCCGCGTTGGCTGCGTTAACGAATGACCATCTGGTGGCGTTGGCATGTCTTGGTGGACGACCCGCGCTCGATGCAGTCAAAAAGGGTCTGCCTCATGCTCCCGCATTGATCAAAAGAACCAACCGGCGGATTCCCGAGAGAACTTCCCATCGAGTCGCGGGATCCCAGCTGGTGAAGAGCGAGCTGGAGGAGAAGAAGTCCGAGCTGCGGCACAAGCTGAAGTACGTGCCCCACGAGTACATCGAGCTGATCGAGATCGCCAGGAACAGCACCCAGGACCGCATCCTGGAGATGAAGGTGATGGAGTTCTTCATGAAGGTGTACGGCTACAGGGGAAAGCACCTGGGCGGAAGCAGAAAGCCTGACGGCGCCATCTATACAGTGGGCAGCCCCATCGATTACGGCGTGATCGTGGACACAAAGGCCTACAGCGGCGGCTACAATCTGCCTATCGGCCAGGCCGACGAGATGCAGAGATACGTGGAGGAGAACCAGACCCGGAATAAGCACATCAACCCCAACGAGTGGTGGAAGGTGTACCCTAGCAGCGTGACCGAGTTCAAGTTCCTGTTCGTGAGCGGCCACTTCAAGGGCAACTACAAGGCCCAGCTGACCAGGCTGAACCACATCACCAACTGCAATGGCGCCGTGCTGAGCGTGGAGGAGCTGCTGATCGGCGGCGAGATGATCAAAGCCGGCACCCTGACACTGGAGGAGGTGCGGCGCAAGTTCAACAACGGCGAGATCAACTTCAGATCT
MDYKDHDGDYKDHDIDYKDDDDKMAPKKKRKVGIHRGVPMVDLRTLGYSQQQQEKIKPKVRSTVAQHHEALVGHGFTHAHIVALSQHPAALGTVAVKYQDMIAALPEATHEAIVGVGKQWSGARALEALLTVAGELRGPPLQLDTGQLLKIAKRGGVTAVEAVHAWRNALTGAPLNLTPDQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPAQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPDQVVAIANNNGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPDQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPAQVVAIANNNGGKQALETVQRLLPVLCQDHGLTPDQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPAQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGRPALESIVAQLSRPDPALAALTNDHLVALACLGGSQLVKSELEEKKSELRHKLKYVPHEYIELIEIARNSTQDRILEMKVMEFFMKVYGYRGKHLGGSRKPDGAIYTVGSPIDYGVIVDTKAYSGGYNLPIGQADEMQRYVEENQTRNKHINPNEWWKVYPSSVTEFKFLFVSGHFKGNYKAQLTRLNHITNCNGAVLSVEELLIGGEMIKAGTLTLEEVRRKFNNGEINFRS
ATGGACTACAAAGACCATGACGGTGATTATAAAGATCATGACATCGATTACAAGGATGACGATGACAAGATGGCCCCCAAGAAGAAGAGGAAGGTGGGCATTCACCGCGGGGTACCTATGGTGGACTTGAGGACACTCGGTTATTCGCAACAGCAACAGGAGAAAATCAAGCCTAAGGTCAGGAGCACCGTCGCGCAACACCACGAGGCGCTTGTGGGGCATGGCTTCACTCATGCGCATATTGTCGCGCTTTCACAGCACCCTGCGGCGCTTGGGACGGTGGCTGTCAAATACCAAGATATGATTGCGGCCCTGCCCGAAGCCACGCACGAGGCAATTGTAGGGGTCGGTAAACAGTGGTCGGGAGCGCGAGCACTTGAGGCGCTGCTGACTGTGGCGGGTGAGCTTAGGGGGCCTCCGCTCCAGCTCGACACCGGGCAGCTGCTGAAGATCGCGAAGAGAGGGGGAGTAACAGCGGTAGAGGCAGTGCACGCCTGGCGCAATGCGCTCACCGGGGCCCCCTTGAACCTGACCCCAGACCAGGTAGTCGCAATCGCGTCGCATGACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCAAACGGAGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCAATGGGGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCCACGACGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGACTCACCCCAGACCAGGTAGTCGCAATCGCGTCGCATGACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCAAATATCGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAATAACAATGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCAATATTGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGGCTGACCCCAGACCAGGTAGTCGCAATCGCGTCGAACATTGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCAAACGGAGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCAATGGGGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCAACAACAACGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGTTTGACCCCAGACCAGGTAGTCGCAATCGCGTCGAACATTGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCAAACGGAGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCAACATCGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCCACGACGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGCCTGACACCCGAACAGGTGGTCGCCATTGCTAGCAACGGGGGAGGACGGCCAGCCTTGGAGTCCATCGTAGCCCAATTGTCCAGGCCCGATCCCGCGTTGGCTGCGTTAACGAATGACCATCTGGTGGCGTTGGCATGTCTTGGTGGATCCCAGCTGGTGAAGAGCGAGCTGGAGGAGAAGAAGTCCGAGCTGCGGCACAAGCTGAAGTACGTGCCCCACGAGTACATCGAGCTGATCGAGATCGCCAGGAACAGCACCCAGGACCGCATCCTGGAGATGAAGGTGATGGAGTTCTTCATGAAGGTGTACGGCTACAGGGGAAAGCACCTGGGCGGAAGCAGAAAGCCTGACGGCGCCATCTATACAGTGGGCAGCCCCATCGATTACGGCGTGATCGTGGACACAAAGGCCTACAGCGGCGGCTACAATCTGCCTATCGGCCAGGCCGACGAGATGCAGAGATACGTGGAGGAGAACCAGACCCGGAATAAGCACATCAACCCCAACGAGTGGTGGAAGGTGTACCCTAGCAGCGTGACCGAGTTCAAGTTCCTGTTCGTGAGCGGCCACTTCAAGGGCAACTACAAGGCCCAGCTGACCAGGCTGAACCACATCACCAACTGCAATGGCGCCGTGCTGAGCGTGGAGGAGCTGCTGATCGGCGGCGAGATGATCAAAGCCGGCACCCTGACACTGGAGGAGGTGCGGCGCAAGTTCAACAACGGCGAGATCAACTTCAGATCT
MDYKDHDGDYKDHDIDYKDDDDKMAPKKKRKVGIHRGVPMVDLRTLGYSQQQQEKIKPKVRSTVAQHHEALVGHGFTHAHIVALSQHPAALGTVAVKYQDMIAALPEATHEAIVGVGKQWSGARALEALLTVAGELRGPPLQLDTGQLLKIAKRGGVTAVEAVHAWRNALTGAPLNLTPDQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPAQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPDQVVAIANNNGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPDQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPAQVVAIANNNGGKQALETVQRLLPVLCQDHGLTPDQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPAQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGRPALESIVAQLSRPDPALAALTNDHLVALACLGGRPALDAVKKGLPHAPALIKRTNRRIPERTSHRVAGSQLVKSELEEKKSELRHKLKYVPHEYIELIEIARNSTQDRILEMKVMEFFMKVYGYRGKHLGGSRKPDGAIYTVGSPIDYGVIVDTKAYSGGYNLPIGQADEMQRYVEENQTRNKHINPNEWWKVYPSSVTEFKFLFVSGHFKGNYKAQLTRLNHITNCNGAVLSVEELLIGGEMIKAGTLTLEEVRRKFNNGEINFRS
ATGGACTACAAAGACCATGACGGTGATTATAAAGATCATGACATCGATTACAAGGATGACGATGACAAGATGGCCCCCAAGAAGAAGAGGAAGGTGGGCATTCACCGCGGGGTACCTATGGTGGACTTGAGGACACTCGGTTATTCGCAACAGCAACAGGAGAAAATCAAGCCTAAGGTCAGGAGCACCGTCGCGCAACACCACGAGGCGCTTGTGGGGCATGGCTTCACTCATGCGCATATTGTCGCGCTTTCACAGCACCCTGCGGCGCTTGGGACGGTGGCTGTCAAATACCAAGATATGATTGCGGCCCTGCCCGAAGCCACGCACGAGGCAATTGTAGGGGTCGGTAAACAGTGGTCGGGAGCGCGAGCACTTGAGGCGCTGCTGACTGTGGCGGGTGAGCTTAGGGGGCCTCCGCTCCAGCTCGACACCGGGCAGCTGCTGAAGATCGCGAAGAGAGGGGGAGTAACAGCGGTAGAGGCAGTGCACGCCTGGCGCAATGCGCTCACCGGGGCCCCCTTGAACCTGACCCCAGACCAGGTAGTCGCAATCGCGTCGCATGACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCAAACGGAGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCAATGGGGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCCACGACGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGACTCACCCCAGACCAGGTAGTCGCAATCGCGTCGCATGACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCAAATATCGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAATAACAATGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCAATATTGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGGCTGACCCCAGACCAGGTAGTCGCAATCGCGTCGAACATTGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCAAACGGAGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCAATGGGGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCAACAACAACGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGTTTGACCCCAGACCAGGTAGTCGCAATCGCGTCGAACATTGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCAAACGGAGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCAACATCGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCCACGACGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGCCTGACACCCGAACAGGTGGTCGCCATTGCTAGCAACGGGGGAGGACGGCCAGCCTTGGAGTCCATCGTAGCCCAATTGTCCAGGCCCGATCCCGCGTTGGCTGCGTTAACGAATGACCATCTGGTGGCGTTGGCATGTCTTGGTGGACGACCCGCGCTCGATGCAGTCAAAAAGGGTCTGCCTCATGCTCCCGCATTGATCAAAAGAACCAACCGGCGGATTCCCGAGAGAACTTCCCATCGAGTCGCGGGATCCCAGCTGGTGAAGAGCGAGCTGGAGGAGAAGAAGTCCGAGCTGCGGCACAAGCTGAAGTACGTGCCCCACGAGTACATCGAGCTGATCGAGATCGCCAGGAACAGCACCCAGGACCGCATCCTGGAGATGAAGGTGATGGAGTTCTTCATGAAGGTGTACGGCTACAGGGGAAAGCACCTGGGCGGAAGCAGAAAGCCTGACGGCGCCATCTATACAGTGGGCAGCCCCATCGATTACGGCGTGATCGTGGACACAAAGGCCTACAGCGGCGGCTACAATCTGCCTATCGGCCAGGCCGACGAGATGCAGAGATACGTGGAGGAGAACCAGACCCGGAATAAGCACATCAACCCCAACGAGTGGTGGAAGGTGTACCCTAGCAGCGTGACCGAGTTCAAGTTCCTGTTCGTGAGCGGCCACTTCAAGGGCAACTACAAGGCCCAGCTGACCAGGCTGAACCACATCACCAACTGCAATGGCGCCGTGCTGAGCGTGGAGGAGCTGCTGATCGGCGGCGAGATGATCAAAGCCGGCACCCTGACACTGGAGGAGGTGCGGCGCAAGTTCAACAACGGCGAGATCAACTTCAGATCT
MDYKDHDGDYKDHDIDYKDDDDKMAPKKKRKVGIHRGVPMVDLRTLGYSQQQQEKIKPKVRSTVAQHHEALVGHGFTHAHIVALSQHPAALGTVAVKYQDMIAALPEATHEAIVGVGKQWSGARALEALLTVAGELRGPPLQLDTGQLLKIAKRGGVTAVEAVHAWRNALTGAPLNLTPDQVVAIASNGGGKQALETVQRLLPVLCQDHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNGGGKQALETVQRLLPVLCQDHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPAQVVAIANNNGGKQALETVQRLLPVLCQDHGLTPDQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNGGGKQALETVQRLLPVLCQDHGLTPDQVVAIANNNGGKQALETVQRLLPVLCQDHGLTPAQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGRPALESIVAQLSRPDPALAALTNDHLVALACLGGSQLVKSELEEKKSELRHKLKYVPHEYIELIEIARNSTQDRILEMKVMEFFMKVYGYRGKHLGGSRKPDGAIYTVGSPIDYGVIVDTKAYSGGYNLPIGQADEMQRYVEENQTRNKHINPNEWWKVYPSSVTEFKFLFVSGHFKGNYKAQLTRLNHITNCNGAVLSVEELLIGGEMIKAGTLTLEEVRRKFNNGEINFRS
ATGGACTACAAAGACCATGACGGTGATTATAAAGATCATGACATCGATTACAAGGATGACGATGACAAGATGGCCCCCAAGAAGAAGAGGAAGGTGGGCATTCACCGCGGGGTACCTATGGTGGACTTGAGGACACTCGGTTATTCGCAACAGCAACAGGAGAAAATCAAGCCTAAGGTCAGGAGCACCGTCGCGCAACACCACGAGGCGCTTGTGGGGCATGGCTTCACTCATGCGCATATTGTCGCGCTTTCACAGCACCCTGCGGCGCTTGGGACGGTGGCTGTCAAATACCAAGATATGATTGCGGCCCTGCCCGAAGCCACGCACGAGGCAATTGTAGGGGTCGGTAAACAGTGGTCGGGAGCGCGAGCACTTGAGGCGCTGCTGACTGTGGCGGGTGAGCTTAGGGGGCCTCCGCTCCAGCTCGACACCGGGCAGCTGCTGAAGATCGCGAAGAGAGGGGGAGTAACAGCGGTAGAGGCAGTGCACGCCTGGCGCAATGCGCTCACCGGGGCCCCCTTGAACCTGACCCCAGACCAGGTAGTCGCAATCGCGTCGAATGGCGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCACATGACGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCAATGGGGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCAACGGTGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGACTCACCCCAGACCAGGTAGTCGCAATCGCGTCGCATGACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCACATGACGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCAACATCGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCAACAACAACGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGGCTGACCCCAGACCAGGTAGTCGCAATCGCGTCGAACATTGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCAAATATCGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCAATGGGGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCAACGGTGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGTTTGACCCCAGACCAGGTAGTCGCAATCGCCAACAATAACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACGCCTGCACAAGTGGTCGCCATCGCCTCCAATATTGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGCCTGACACCCGAACAGGTGGTCGCCATTGCTAGCAACGGGGGAGGACGGCCAGCCTTGGAGTCCATCGTAGCCCAATTGTCCAGGCCCGATCCCGCGTTGGCTGCGTTAACGAATGACCATCTGGTGGCGTTGGCATGTCTTGGTGGATCCCAGCTGGTGAAGAGCGAGCTGGAGGAGAAGAAGTCCGAGCTGCGGCACAAGCTGAAGTACGTGCCCCACGAGTACATCGAGCTGATCGAGATCGCCAGGAACAGCACCCAGGACCGCATCCTGGAGATGAAGGTGATGGAGTTCTTCATGAAGGTGTACGGCTACAGGGGAAAGCACCTGGGCGGAAGCAGAAAGCCTGACGGCGCCATCTATACAGTGGGCAGCCCCATCGATTACGGCGTGATCGTGGACACAAAGGCCTACAGCGGCGGCTACAATCTGCCTATCGGCCAGGCCGACGAGATGCAGAGATACGTGGAGGAGAACCAGACCCGGAATAAGCACATCAACCCCAACGAGTGGTGGAAGGTGTACCCTAGCAGCGTGACCGAGTTCAAGTTCCTGTTCGTGAGCGGCCACTTCAAGGGCAACTACAAGGCCCAGCTGACCAGGCTGAACCACATCACCAACTGCAATGGCGCCGTGCTGAGCGTGGAGGAGCTGCTGATCGGCGGCGAGATGATCAAAGCCGGCACCCTGACACTGGAGGAGGTGCGGCGCAAGTTCAACAACGGCGAGATCAACTTCAGATCT
MDYKDHDGDYKDHDIDYKDDDDKMAPKKKRKVGIHRGVPMVDLRTLGYSQQQQEKIKPKVRSTVAQHHEALVGHGFTHAHIVALSQHPAALGTVAVKYQDMIAALPEATHEAIVGVGKQWSGARALEALLTVAGELRGPPLQLDTGQLLKIAKRGGVTAVEAVHAWRNALTGAPLNLTPDQVVAIASNGGGKQALETVQRLLPVLCQDHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNGGGKQALETVQRLLPVLCQDHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPAQVVAIANNNGGKQALETVQRLLPVLCQDHGLTPDQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNGGGKQALETVQRLLPVLCQDHGLTPDQVVAIANNNGGKQALETVQRLLPVLCQDHGLTPAQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGRPALESIVAQLSRPDPALAALTNDHLVALACLGGRPALDAVKKGLPHAPALIKRTNRRIPERTSHRVAGSQLVKSELEEKKSELRHKLKYVPHEYIELIEIARNSTQDRILEMKVMEFFMKVYGYRGKHLGGSRKPDGAIYTVGSPIDYGVIVDTKAYSGGYNLPIGQADEMQRYVEENQTRNKHINPNEWWKVYPSSVTEFKFLFVSGHFKGNYKAQLTRLNHITNCNGAVLSVEELLIGGEMIKAGTLTLEEVRRKFNNGEINFRS
ATGGACTACAAAGACCATGACGGTGATTATAAAGATCATGACATCGATTACAAGGATGACGATGACAAGATGGCCCCCAAGAAGAAGAGGAAGGTGGGCATTCACCGCGGGGTACCTATGGTGGACTTGAGGACACTCGGTTATTCGCAACAGCAACAGGAGAAAATCAAGCCTAAGGTCAGGAGCACCGTCGCGCAACACCACGAGGCGCTTGTGGGGCATGGCTTCACTCATGCGCATATTGTCGCGCTTTCACAGCACCCTGCGGCGCTTGGGACGGTGGCTGTCAAATACCAAGATATGATTGCGGCCCTGCCCGAAGCCACGCACGAGGCAATTGTAGGGGTCGGTAAACAGTGGTCGGGAGCGCGAGCACTTGAGGCGCTGCTGACTGTGGCGGGTGAGCTTAGGGGGCCTCCGCTCCAGCTCGACACCGGGCAGCTGCTGAAGATCGCGAAGAGAGGGGGAGTAACAGCGGTAGAGGCAGTGCACGCCTGGCGCAATGCGCTCACCGGGGCCCCCTTGAACCTGACCCCAGACCAGGTAGTCGCAATCGCGTCGAATGGCGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCACATGACGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCAATGGGGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCAACGGTGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGACTCACCCCAGACCAGGTAGTCGCAATCGCGTCGCATGACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCACATGACGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCAACATCGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCAACAACAACGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGGCTGACCCCAGACCAGGTAGTCGCAATCGCGTCGAACATTGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACACCGGAGCAAGTCGTGGCCATTGCATCAAATATCGGTGGCAAACAGGCTCTTGAGACGGTTCAGAGACTTCTCCCAGTTCTCTGTCAAGCCCACGGGCTGACTCCCGATCAAGTTGTAGCGATTGCGAGCAATGGGGGAGGGAAACAAGCATTGGAGACTGTCCAACGGCTCCTTCCCGTGTTGTGTCAAGCCCACGGTTTGACGCCTGCACAAGTGGTCGCCATCGCCTCCAACGGTGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGTTTGACCCCAGACCAGGTAGTCGCAATCGCCAACAATAACGGGGGAAAGCAAGCCCTGGAAACCGTGCAAAGGTTGTTGCCGGTCCTTTGTCAAGACCACGGCCTTACGCCTGCACAAGTGGTCGCCATCGCCTCCAATATTGGCGGTAAGCAGGCGCTGGAAACAGTACAGCGCCTGCTGCCTGTACTGTGCCAGGATCATGGCCTGACACCCGAACAGGTGGTCGCCATTGCTAGCAACGGGGGAGGACGGCCAGCCTTGGAGTCCATCGTAGCCCAATTGTCCAGGCCCGATCCCGCGTTGGCTGCGTTAACGAATGACCATCTGGTGGCGTTGGCATGTCTTGGTGGACGACCCGCGCTCGATGCAGTCAAAAAGGGTCTGCCTCATGCTCCCGCATTGATCAAAAGAACCAACCGGCGGATTCCCGAGAGAACTTCCCATCGAGTCGCGGGATCCCAGCTGGTGAAGAGCGAGCTGGAGGAGAAGAAGTCCGAGCTGCGGCACAAGCTGAAGTACGTGCCCCACGAGTACATCGAGCTGATCGAGATCGCCAGGAACAGCACCCAGGACCGCATCCTGGAGATGAAGGTGATGGAGTTCTTCATGAAGGTGTACGGCTACAGGGGAAAGCACCTGGGCGGAAGCAGAAAGCCTGACGGCGCCATCTATACAGTGGGCAGCCCCATCGATTACGGCGTGATCGTGGACACAAAGGCCTACAGCGGCGGCTACAATCTGCCTATCGGCCAGGCCGACGAGATGCAGAGATACGTGGAGGAGAACCAGACCCGGAATAAGCACATCAACCCCAACGAGTGGTGGAAGGTGTACCCTAGCAGCGTGACCGAGTTCAAGTTCCTGTTCGTGAGCGGCCACTTCAAGGGCAACTACAAGGCCCAGCTGACCAGGCTGAACCACATCACCAACTGCAATGGCGCCGTGCTGAGCGTGGAGGAGCTGCTGATCGGCGGCGAGATGATCAAAGCCGGCACCCTGACACTGGAGGAGGTGCGGCGCAAGTTCAACAACGGCGAGATCAACTTCAGATCT
5’AGCGCCCAATACGCAAACCGCCTCTCCCCGCGCGTTGGCCGATTCATTAATGCAGCTGGCACGACAGGTTTCCCGACTGGAAAGCGGGCAGTGAGCGCAACGCAATTAATGTGAGTTAGCTCACTCATTAGGCACCCCAGGCTTTACACTTTATGCTTCCGGCTCGTATGTTGTGTGGAATTGTGAGCGGATAACAATTTCACACAGGAAACAGCTATGACCATGATTACGCCAAGCTCAGAATTAACCCTCACTAAAGGGACTAGTCCTGCAGGTTTAAACGAATTCGCCCTTGATACTTATTAACCATACCTTGGAGGGGAAATCACACATGAAAAGTGTCATTTCTTTACTAATCATATTCATGTCTTTTCTCCCCATAGCAAGACAAAGACCTGTTTTAAACACATTTACAACCTATATGTTGCCTTGTACTAGGTAAAAAGTTGTACATTTCTGAAATAATTTTGGTATTTCTGTTCAGATCACTAAACTCAAGAATCAGCAATTCTCTGAGGCTTTCTTTTAAATATACATAAGGAACTTTCGGAGTGAAGGGAGAGTTTGTCAATAACTTGATGCATGTGAAGGGGAGATAAAAAGGTTGCTATTTTTCATCAACATATTTTGATTTGGCTTTCTATAATTGATGGGCTTAAAAGATCTAATCTACTTTAAACAGATGCCAAATAAATGGATGAATCTTAGACCCTCTATAACAGTAACTTCCTTTTAAAAAAGACCTCTCCCACCCCACCCCCAGCCCAGGCTGTGTATGAAAACTAAGCCATGTGCACAACTCTGACTGGGTCACCAGCCCACTTGAGTCCGTGTCACAAGCCCACAGATATTTCCTGCTCCCCAGTGGATCGGGTGTAAACTGAGCTTGCTCGCTCGGGAGCCTCTTGCTGGAAAATAGAACAGCATTTGCAGAAGCGTTTGGCAATGTGCTTTTGGAAGAAGACTAAGAGGTAGTTTCTGAACTTCTCCCCGACAAAGGCATAGATGATGGGGTTGATGCAGCAGTGCGTCATCCCAAGAGTCTCTGTCACCTGCATAGCTTGGTCCAACCTGTTAGAGCTACTGCAATTATTCAGGCCAAAGAATTCCTGGAAGGTGTTCAGGAGAAGGACAATGTTGTAGGGAGCCCAGAAGAGAAAATAAACAATCATGATGGTGAAGATAAGCCTCACAGCCCTGTGCCTCTTCTTCTCATTTCGACACCGAAGCAGAGTTTTTAGGATTCCCGAGTAGCAGATGACCATGACAAGCAGCGGCAGGACCAGCCCCAAGATGACTATCTTTAATGTCTGGAAATTCTTCCAGAATTGATACTGACTGTATGGAAAATGAGAGCTGCAGGTGTAATGAAGACCTTCTTTTTGAGATCTGGTAAAGATGATTCCTGGGAGAGACGCAAACACAGCCACCACCCAAGTGATCACACTTGTCACCACCCCAAAGGTGACCGTCCTGGCTTTTAAAGCAAACACAGCATGGACGACAGCCAGGTACCTATCGATTGTCAGGAGGATGATGAAGAAGATTCCAGAGAAGAAGCCTATAAAATAGAGCCCTGTCAAGAGTTGACACATTGTATTTCCAAAGTCCCACTGGGCGGCAGCATAGTGAGCCCAGAAGGGGACAGTAAGAAGGAAAAACAGGTCAGAGATGGCCAGGTTGAGCAGGTAGATGTCAGTCATGCTCTTCAGCCTTTTGCAGTTTTCTAGACGAGGCATCCAGTCCAGACGCCATCAGGGCATACTCACTGATCTAGATGAGGATGACCAGCATGTTGCCCACAAAACCAAAGATGAACACCAGTGAGTAGAGCGGAGGCAGGAGGCGGGCTGCGATTTGCTTCACATTGATTTTTTGGCAGGGCTCCGATGTATAATAATTGATGTCATAGATTGGACTTGACACTTGATAATCCATCTTGTTCCACCCTGTGCATAAATAAAAAGTGATCTTTTATAAAGTCCTAGAATGTATTTAGTTGCCCTCCATGAATGCAAACTGTTTTATACATCAATAGGTTTTTAATTGCCTACATAGATGTCTACATTGAATTAACTCTCTTTTTGGCCAAGCAATGAAGTTTTGTAGTGAAGGGAAGGTTTGCTGCTAGCTTCCCTGTCCACTAGATGGAGAGCTTGGCTCTGTTGGGGGAATTCATGAAAGCACCATCTCACCAAATAAAATCTTGTGCTCTATAGCACCATGGAGTGAATGAAGCTTTGACAACAATTAAGGGCGAATTCGCGGCCGCTAAATTCAATTCGCCCTATAGTGAGTCGTATTACAATTCACTGGCCGTCGTTTTACAACGTCGTGACTGGGAAAACCCTGGCGTTACCCAACTTAATCGCCTTGCAGCACATCCCCCTTTCGCCAGCTGGCGTAATAGCGAAGAGGCCCGCACCGATCGCCCTTCCCAACAGTTGCGCAGCCTATACGTACGGCAGTTTAAGGTTTACACCTATAAAAGAGAGAGCCGTTATCGTCTGTTTGTGGATGTACAGAGTGATATTATTGACACGCCGGGGCGACGGATGGTGATCCCCCTGGCCAGTGCACGTCTGCTGTCAGATAAAGTCTCCCGTGAACTTTACCCGGTGGTGCATATCGGGGATGAAAGCTGGCGCATGATGACCACCGATATGGCCAGTGTGCCGGTCTCCGTTATCGGGGAAGAAGTGGCTGATCTCAGCCACCGCGAAAATGACATCAAAAACGCCATTAACCTGATGTTCTGGGGAATATAAATGTCAGGCATGAGATTATCAAAAAGGATCTTCACCTAGATCCTTTTCACGTAGAAAGCCAGTCCGCAGAAACGGTGCTGACCCCGGATGAATGTCAGCTACTGGGCTATCTGGACAAGGGAAAACGCAAGCGCAAAGAGAAAGCAGGTAGCTTGCAGTGGGCTTACATGGCGATAGCTAGACTGGGCGGTTTTATGGACAGCAAGCGAACCGGAATTGCCAGCTGGGGCGCCCTCTGGTAAGGTTGGGAAGCCCTGCAAAGTAAACTGGATGGCTTTCTTGCCGCCAAGGATCTGATGGCGCAGGGGATCAAGCTCTGATCAAGAGACAGGATGAGGATCGTTTCGCATGATTGAACAAGATGGATTGCACGCAGGTTCTCCGGCCGCTTGGGTGGAGAGGCTATTCGGCTATGACTGGGCACAACAGACAATCGGCTGCTCTGATGCCGCCGTGTTCCGGCTGTCAGCGCAGGGGCGCCCGGTTCTTTTTGTCAAGACCGACCTGTCCGGTGCCCTGAATGAACTGCAAGACGAGGCAGCGCGGCTATCGTGGCTGGCCACGACGGGCGTTCCTTGCGCAGCTGTGCTCGACGTTGTCACTGAAGCGGGAAGGGACTGGCTGCTATTGGGCGAAGTGCCGGGGCAGGATCTCCTGTCATCTCACCTTGCTCCTGCCGAGAAAGTATCCATCATGGCTGATGCAATGCGGCGGCTGCATACGCTTGATCCGGCTACCTGCCCATTCGACCACCAAGCGAAACATCGCATCGAGCGAGCACGTACTCGGATGGAAGCCGGTCTTGTCGATCAGGATGATCTGGACGAAGAGCATCAGGGGCTCGCGCCAGCCGAACTGTTCGCCAGGCTCAAGGCGAGCATGCCCGACGGCGAGGATCTCGTCGTGACCCATGGCGATGCCTGCTTGCCGAATATCATGGTGGAAAATGGCCGCTTTTCTGGATTCATCGACTGTGGCCGGCTGGGTGTGGCGGACCGCTATCAGGACATAGCGTTGGCTACCCGTGATATTGCTGAAGAGCTTGGCGGCGAATGGGCTGACCGCTTCCTCGTGCTTTACGGTATCGCCGCTCCCGATTCGCAGCGCATCGCCTTCTATCGCCTTCTTGACGAGTTCTTCTGAATTATTAACGCTTACAATTTCCTGATGCGGTATTTTCTCCTTACGCATCTGTGCGGTATTTCACACCGCATCAGGTGGCACTTTTCGGGGAAATGTGCGCGGAACCCCTATTTGTTTATTTTTCTAAATACATTCAAATATGTATCCGCTCATGAGATTATCAAAAAGGATCTTCACCTAGATCCTTTTAAATTAAAAATGAAGTTTTAAATCAATCTAAAGTATATATGAGTAAACTTGGTCTGACAGTTACCAATGCTTAATCAGTGAGGCACCTATCTCAGCGATCTGTCTATTTCGTTCATCCATAGTTGCCTGACTCCCCGTCGTGTAGATAACTACGATACGGGAGGGCTTACCATCTGGCCCCAGTGCTGCAATGATACCGCGAGACCCACGCTCACCGGCTCCAGATTTATCAGCAATAAACCAGCCAGCCGGAAGGGCCGAGCGCAGAAGTGGTCCTGCAACTTTATCCGCCTCCATCCAGTCTATTAATTGTTGCCGGGAAGCTAGAGTAAGTAGTTCGCCAGTTAATAGTTTGCGCAACGTTGTTGCCATTGCTACAGGCATCGTGGTGTCACGCTCGTCGTTTGGTATGGCTTCATTCAGCTCCGGTTCCCAACGATCAAGGCGAGTTACATGATCCCCCATGTTGTGCAAAAAAGCGGTTAGCTCCTTCGGTCCTCCGATCGTTGTCAGAAGTAAGTTGGCCGCAGTGTTATCACTCATGGTTATGGCAGCACTGCATAATTCTCTTACTGTCATGCCATCCGTAAGATGCTTTTCTGTGACTGGTGAGTACTCAACCAAGTCATTCTGAGAATAGTGTATGCGGCGACCGAGTTGCTCTTGCCCGGCGTCAATACGGGATAATACCGCGCCACATAGCAGAACTTTAAAAGTGCTCATCATTGGAAAACGTTCTTCGGGGCGAAAACTCTCAAGGATCTTACCGCTGTTGAGATCCAGTTCGATGTAACCCACTCGTGCACCCAACTGATCTTCAGCATCTTTTACTTTCACCAGCGTTTCTGGGTGAGCAAAAACAGGAAGGCAAAATGCCGCAAAAAAGGGAATAAGGGCGACACGGAAATGTTGAATACTCATACTCTTCCTTTTTCAATATTATTGAAGCATTTATCAGGGTTATTGTCTCATGACCAAAATCCCTTAACGTGAGTTTTCGTTCCACTGAGCGTCAGACCCCGTAGAAAAGATCAAAGGATCTTCTTGAGATCCTTTTTTTCTGCGCGTAATCTGCTGCTTGCAAACAAAAAAACCACCGCTACCAGCGGTGGTTTGTTTGCCGGATCAAGAGCTACCAACTCTTTTTCCGAAGGTAACTGGCTTCAGCAGAGCGCAGATACCAAATACTGTTCTTCTAGTGTAGCCGTAGTTAGGCCACCACTTCAAGAACTCTGTAGCACCGCCTACATACCTCGCTCTGCTAATCCTGTTACCAGTGGCTGCTGCCAGTGGCGATAAGTCGTGTCTTACCGGGTTGGACTCAAGACGATAGTTACCGGATAAGGCGCAGCGGTCGGGCTGAACGGGGGGTTCGTGCACACAGCCCAGCTTGGAGCGAACGACCTACACCGAACTGAGATACCTACAGCGTGAGCTATGAGAAAGCGCCACGCTTCCCGAAGGGAGAAAGGCGGACAGGTATCCGGTAAGCGGCAGGGTCGGAACAGGAGAGCGCACGAGGGAGCTTCCAGGGGGAAACGCCTGGTATCTTTATAGTCCTGTCGGGTTTCGCCACCTCTGACTTGAGCGTCGATTTTTGTGATGCTCGTCAGGGGGGCGGAGCCTATGGAAAAACGCCAGCAACGCGGCCTTTTTACGGTTCCTGGCCTTTTGCTGGCCTTTTGCTCACATGTTCTTTCCTGCGTTATCCCCTGATTCTGTGGATAACCGTATTACCGCCTTTGAGTGAGCTGATACCGCTCGCCGCAGCCGAACGACCGAGCGCAGCGAGTCAGTGAGCGAGGAAGCGGAAG3’(配列番号176)
コード配列に下線を引いた、完全なTALE構築物配列(配列番号303):
TAATACGACTCACTATAGGGAGACCCAAGCTGGCTAGCTTAAGCTGATCCACTAGTCCAGTGTGGTGGAATTCGCTAGCGCCACCATGGCCCCCAAGAAGAAGAGGAAGGTGGGAATCGATGGGGTACCCGCCGCTGTGGATCTACGCACGCTCGGCTACAGCCAGCAGCAACAGGAGAAGATCAAACCGAAGGTTCGTTCGACAGTGGCGCAGCACCACGAGGCACTGGTCGGCCATGGGTTTACACACGCGCACATCGTTGCGCTCAGCCAACACCCGGCAGCGTTAGGGACCGTCGCTGTCAAGTATCAGGACATGATCGCAGCGTTGCCAGAGGCGACACACGAAGCGATCGTTGGCGTCGGCAAACAGTGGTCCGGCGCACGCGCCCTGGAGGCCTTGCTCACGGTGGCGGGAGAGTTGAGAGGTCCACCGTTACAGTTGGACACAGGCCAACTTCTCAAGATTGCAAAACGTGGCGGCGTGACCGCAGTGGAGGCAGTGCATGCATGGCGCAATGCACTGACGGGGGCCCCCCTGAACCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCAATATTGGTGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCAATGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCAATATTGGTGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCAATATTGGTGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGCACAGGTGGTGGCCATCGCCAGCAATATTGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGCGCCAGGCCCATGGCCTGACCCCGGAGCAGGTCGTGGCCATCGCCAGCAATGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGCACAGGTGGTGGCCATCGCCAGCAATATTGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTCGTGGCCATCGCCAGCCACGATGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCCACGATGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGACCAGGTGGTGGCCATCGCCAGCAATGGCGGTGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGACCAGGTGGTGGCCATCGCCAGCAATGGCGGTGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCCACGATGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCAATGGCGGCGGCAGGCCGGCGCTGGAGAGCATTGTTGCCCAGTTATCTCGCCCTGATCCGGCGTTGGCCGCGTTAACCAACGACCACCTCGTCGCCTTGGCCTGCCTCGGCGGACGTCCTGCGCTGGATGCAGTGAAAAAGGGATTGCCGCACGCGCCGGCCTTGATCAAAAGAACCAATCGCCGTATTCCCGAACGCACATCCCATCGCGTTGCCGACCACGCGCAAGTGGTTCGCGTGCTGGGTTTTTTCCAGTGCCACTCCCACCCAGCGCAAGCATTTGATGACGCCATGACGCAGTTCGGGATGAGCAGGCACGGGTTGTTACAGCTCTTTCGCAGAGTGGGCGTCACCGAACTCGAAGCCCGCAGTGGAACGCTCCCCCCAGCCTCGCAGCGTTGGGACCGTATCCTCCAGGCATCAGGGATGAAAAGGGCCAAACCGTCCCCTACTTCAACTCAAACGCCGGACCAGGCGTCTTTGCATGCATTCGCCGATTCGCTGGAGCGTGACCTTGATGCGCCCAGCCCAACGCACGAGGGAGATCAGAGGCGGGCAAGCAGCCGTAAACGGTCCCGATCGGATCGTGCTGTCACCGGTCCCTCCGCACAGCAATCGTTCGAGGTGCGCGCTCCCGAACAGCGCGATGCGCTGCATTTGCCCCTCAGTTGGAGGGTAAAACGCCCGCGTACCAGTATCGGGGGCGGCCTCCCGGATCCTGGTACGCCCACGGCTGCCGACCTGGCAGCGTCCAGCACCGTGATGCGGGAACAAGATGAGGACCCCTTCGCAGGGGCAGCGGATGATTTCCCGGCATTCAACGAAGAGGAGCTCGCATGGTTGATGGAGCTATTGCCTCAGGACCGCGGCCGCGCCCCCCCGACCGATGTCAGCCTGGGGGACGAGCTCCACTTAGACGGCGAGGACGTGGCGATGGCGCATGCCGACGCGCTAGACGATTTCGATCTGGACATGTTGGGGGACGGGGATTCCCCGGGTCCGGGATTTACCCCCCACGACTCCGCCCCCTACGGCGCTCTGGATATGGCCGACTTCGAGTTTGAGCAGATGTTTACCGATGCCCTTGGAATTGACGAGTACGGTGGCGGCCGCGACTACAAGGACGACGATGACAAGTAAGCTTCTCGAGTCTAGCTAGTTTAAACCCGCTGATCAGCCTCGACTGTGCCTTCTAGTTGCCAGCCATCTGTTGTTTGCCCCTCCCCCGTGCCTTCCTTGACCCTGGAAGGTGCCACTCCCACTGTCCTTTCCTAATAAAATGAGGAAATTGCATCGCATTGTCTGAGTAGGTGTCATTCTATTCTGGGGGGTGGGGTGGGGCAGGACAGCAAGGGGGAGGATTGGGAAGACAATAGCAGGCATGCTGGGGATGCGGTGGGCTCTATGGCTTCTGAGGCGGAAAGAACCAGCTGGGGCTCTAGGGGGTATCCCCACGCGCCCTGTAGCGGCGCATTAAGCGCGGCGGGTGTGGTGGTTACGCGCAGCGTGACCGCTACACTTGCCAGCGCCCTAGCGCCCGCTCCTTTCGCTTTCTTCCCTTCCTTTCTCGCCACGTTCGCCGGCTTTCCCCGTCAAGCTCTAAATCGGGGCATCCCTTTAGGGTTCCGATTTAGTGCTTTACGGCACCTCGACCCCAAAAAACTTGATTAGGGTGATGGTTCACGTAGTGGGCCATCGCCCTGATAGACGGTTTTTCGCCCTTTGACGTTGGAGTCCACGTTCTTTAATAGTGGACTCTTGTTCCAAACTGGAACAACACTCAACCCTATCTCGGTCTATTCTTTTGATTTATAAGGGATTTTGGGGATTTCGGCCTATTGGTTAAAAAATGAGCTGATTTAACAAAAATTTAACGCGAATTAATTCTGTGGAATGTGTGTCAGTTAGGGTGTGGAAAGTCCCCAGGCTCCCCAGGCAGGCAGAAGTATGCAAAGCATGCATCTCAATTAGTCAGCAACCAGGTGTGGAAAGTCCCCAGGCTCCCCAGCAGGCAGAAGTATGCAAAGCATGCATCTCAATTAGTCAGCAACCATAGTCCCGCCCCTAACTCCGCCCATCCCGCCCCTAACTCCGCCCAGTTCCGCCCATTCTCCGCCCCATGGCTGACTAATTTTTTTTATTTATGCAGAGGCCGAGGCCGCCTCTGCCTCTGAGCTATTCCAGAAGTAGTGAGGAGGCTTTTTTGGAGGCCTAGGCTTTTGCAAAAAGCTCCCGGGAGCTTGTATATCCATTTTCGGATCTGATCAAGAGACAGGATGAGGATCGTTTCGCATGATTGAACAAGATGGATTGCACGCAGGTTCTCCGGCCGCTTGGGTGGAGAGGCTATTCGGCTATGACTGGGCACAACAGACAATCGGCTGCTCTGATGCCGCCGTGTTCCGGCTGTCAGCGCAGGGGCGCCCGGTTCTTTTTGTCAAGACCGACCTGTCCGGTGCCCTGAATGAACTGCAGGACGAGGCAGCGCGGCTATCGTGGCTGGCCACGACGGGCGTTCCTTGCGCAGCTGTGCTCGACGTTGTCACTGAAGCGGGAAGGGACTGGCTGCTATTGGGCGAAGTGCCGGGGCAGGATCTCCTGTCATCTCACCTTGCTCCTGCCGAGAAAGTATCCATCATGGCTGATGCAATGCGGCGGCTGCATACGCTTGATCCGGCTACCTGCCCATTCGACCACCAAGCGAAACATCGCATCGAGCGAGCACGTACTCGGATGGAAGCCGGTCTTGTCGATCAGGATGATCTGGACGAAGAGCATCAGGGGCTCGCGCCAGCCGAACTGTTCGCCAGGCTCAAGGCGCGCATGCCCGACGGCGAGGATCTCGTCGTGACCCATGGCGATGCCTGCTTGCCGAATATCATGGTGGAAAATGGCCGCTTTTCTGGATTCATCGACTGTGGCCGGCTGGGTGTGGCGGACCGCTATCAGGACATAGCGTTGGCTACCCGTGATATTGCTGAAGAGCTTGGCGGCGAATGGGCTGACCGCTTCCTCGTGCTTTACGGTATCGCCGCTCCCGATTCGCAGCGCATCGCCTTCTATCGCCTTCTTGACGAGTTCTTCTGAGCGGGACTCTGGGGTTCGAAATGACCGACCAAGCGACGCCCAACCTGCCATCACGAGATTTCGATTCCACCGCCGCCTTCTATGAAAGGTTGGGCTTCGGAATCGTTTTCCGGGACGCCGGCTGGATGATCCTCCAGCGCGGGGATCTCATGCTGGAGTTCTTCGCCCACCCCAACTTGTTTATTGCAGCTTATAATGGTTACAAATAAAGCAATAGCATCACAAATTTCACAAATAAAGCATTTTTTTCACTGCATTCTAGTTGTGGTTTGTCCAAACTCATCAATGTATCTTATCATGTCTGTATACCGTCGACCTCTAGCTAGAGCTTGGCGTAATCATGGTCATAGCTGTTTCCTGTGTGAAATTGTTATCCGCTCACAATTCCACACAACATACGAGCCGGAAGCATAAAGTGTAAAGCCTGGGGTGCCTAATGAGTGAGCTAACTCACATTAATTGCGTTGCGCTCACTGCCCGCTTTCCAGTCGGGAAACCTGTCGTGCCAGCTGCATTAATGAATCGGCCAACGCGCGGGGAGAGGCGGTTTGCGTATTGGGCGCTCTTCCGCTTCCTCGCTCACTGACTCGCTGCGCTCGGTCGTTCGGCTGCGGCGAGCGGTATCAGCTCACTCAAAGGCGGTAATACGGTTATCCACAGAATCAGGGGATAACGCAGGAAAGAACATGTGAGCAAAAGGCCAGCAAAAGGCCAGGAACCGTAAAAAGGCCGCGTTGCTGGCGTTTTTCCATAGGCTCCGCCCCCCTGACGAGCATCACAAAAATCGACGCTCAAGTCAGAGGTGGCGAAACCCGACAGGACTATAAAGATACCAGGCGTTTCCCCCTGGAAGCTCCCTCGTGCGCTCTCCTGTTCCGACCCTGCCGCTTACCGGATACCTGTCCGCCTTTCTCCCTTCGGGAAGCGTGGCGCTTTCTCAATGCTCACGCTGTAGGTATCTCAGTTCGGTGTAGGTCGTTCGCTCCAAGCTGGGCTGTGTGCACGAACCCCCCGTTCAGCCCGACCGCTGCGCCTTATCCGGTAACTATCGTCTTGAGTCCAACCCGGTAAGACACGACTTATCGCCACTGGCAGCAGCCACTGGTAACAGGATTAGCAGAGCGAGGTATGTAGGCGGTGCTACAGAGTTCTTGAAGTGGTGGCCTAACTACGGCTACACTAGAAGGACAGTATTTGGTATCTGCGCTCTGCTGAAGCCAGTTACCTTCGGAAAAAGAGTTGGTAGCTCTTGATCCGGCAAACAAACCACCGCTGGTAGCGGTGGTTTTTTTGTTTGCAAGCAGCAGATTACGCGCAGAAAAAAAGGATCTCAAGAAGATCCTTTGATCTTTTCTACGGGGTCTGACGCTCAGTGGAACGAAAACTCACGTTAAGGGATTTTGGTCATGAGATTATCAAAAAGGATCTTCACCTAGATCCTTTTAAATTAAAAATGAAGTTTTAAATCAATCTAAAGTATATATGAGTAAACTTGGTCTGACAGTTACCAATGCTTAATCAGTGAGGCACCTATCTCAGCGATCTGTCTATTTCGTTCATCCATAGTTGCCTGACTCCCCGTCGTGTAGATAACTACGATACGGGAGGGCTTACCATCTGGCCCCAGTGCTGCAATGATACCGCGAGACCCACGCTCACCGGCTCCAGATTTATCAGCAATAAACCAGCCAGCCGGAAGGGCCGAGCGCAGAAGTGGTCCTGCAACTTTATCCGCCTCCATCCAGTCTATTAATTGTTGCCGGGAAGCTAGAGTAAGTAGTTCGCCAGTTAATAGTTTGCGCAACGTTGTTGCCATTGCTACAGGCATCGTGGTGTCACGCTCGTCGTTTGGTATGGCTTCATTCAGCTCCGGTTCCCAACGATCAAGGCGAGTTACATGATCCCCCATGTTGTGCAAAAAAGCGGTTAGCTCCTTCGGTCCTCCGATCGTTGTCAGAAGTAAGTTGGCCGCAGTGTTATCACTCATGGTTATGGCAGCACTGCATAATTCTCTTACTGTCATGCCATCCGTAAGATGCTTTTCTGTGACTGGTGAGTACTCAACCAAGTCATTCTGAGAATAGTGTATGCGGCGACCGAGTTGCTCTTGCCCGGCGTCAATACGGGATAATACCGCGCCACATAGCAGAACTTTAAAAGTGCTCATCATTGGAAAACGTTCTTCGGGGCGAAAACTCTCAAGGATCTTACCGCTGTTGAGATCCAGTTCGATGTAACCCACTCGTGCACCCAACTGATCTTCAGCATCTTTTACTTTCACCAGCGTTTCTGGGTGAGCAAAAACAGGAAGGCAAAATGCCGCAAAAAAGGGAATAAGGGCGACACGGAAATGTTGAATACTCATACTCTTCCTTTTTCAATATTATTGAAGCATTTATCAGGGTTATTGTCTCATGAGCGGATACATATTTGAATGTATTTAGAAAAATAAACAAATAGGGGTTCCGCGCACATTTCCCCGAAAAGTGCCACCTGACGTCGACGGATCGGGAGATCTCCCGATCCCCTATGGTCGACTCTCAGTACAATCTGCTCTGATGCCGCATAGTTAAGCCAGTATCTGCTCCCTGCTTGTGTGTTGGAGGTCGCTGAGTAGTGCGCGAGCAAAATTTAAGCTACAACAAGGCAAGGCTTGACCGACAATTGCATGAAGAATCTGCTTAGGGTTAGGCGTTTTGCGCTGCTTCGCGATGTACGGGCCAGATATACGCGTTGACATTGATTATTGACTAGTTATTAATAGTAATCAATTACGGGGTCATTAGTTCATAGCCCATATATGGAGTTCCGCGTTACATAACTTACGGTAAATGGCCCGCCTGGCTGACCGCCCAACGACCCCCGCCCATTGACGTCAATAATGACGTATGTTCCCATAGTAACGCCAATAGGGACTTTCCATTGACGTCAATGGGTGGACTATTTACGGTAAACTGCCCACTTGGCAGTACATCAAGTGTATCATATGCCAAGTACGCCCCCTATTGACGTCAATGACGGTAAATGGCCCGCCTGGCATTATGCCCAGTACATGACCTTATGGGACTTTCCTACTTGGCAGTACATCTACGTATTAGTCATCGCTATTACCATGGTGATGCGGTTTTGGCAGTACATCAATGGGCGTGGATAGCGGTTTGACTCACGGGGATTTCCAAGTCTCCACCCCATTGACGTCAATGGGAGTTTGTTTTGGCACCAAAATCAACGGGACTTTCCAAAATGTCGTAACAACTCCGCCCCATTGACGCAAATGGGCGGTAGGCGTGTACGGTGGGAGGTCTATATAAGCAGAGCTCTCTGGCTAACTAGAGAACCCACTGCTTACTGGCTTATCGAAAT
それぞれの発現構築物の配列を再生成するために、上述の構築物の下線を引いた領域を以下に示されるそれぞれのCDSに置換する。
MVDLRTLGYSQQQQEKIKPKVRSTVAQHHEALVGHGFTHAHIVALSQHPAALGTVAVKYQDMIAALPEATHEAIVGVGKQWSGARALEALLTVAGELRGPPLQLDTGQLLKIAKRGGVTAVEAVHAWRNALTGAPLNLTPEQVVAIASNNGGKQALETVQRLLPVLCQAHGLTPQQVVAIASNNGGKQALETVQRLLPVLCQAHGLTPQQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNNGGKQALETVQRLLPVLCQAHGLTPEQVVAIASHDGGKQALETVQALLPVLCQAHGLTPEQVVAIASHDGGKQALETVQALLPVLCQAHGLTPEQVVAIASNIGGKQALETVQALLPVLCQAHGLTPEQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPQQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNNGGKQALETVQALLPVLCQAHGLTPEQVVAIASNKGGKQALETVQRLLPVLCQAHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNNGGKQALETVQRLLPVLCQAHGLTPQQVVAIASNKGGRPALETVQRLLPVLCQAHGLTPEQVVAIASNNGGKQALETVQRLLPVLCQAHGLTPQQVVAIASNGGGRPALESIVAQLSRPDPALAALTNDHLVALACLGGRPALDAVKKGLPHAPALIKRTNRRIPERTSHRVADHAQVVRVLGFFQCHSHPAQAFDDAMTQFGMSRHGLLQLFRRVGVTELEARSGTLPPASQRWDRILQASGMKRAKPSPTSTQTPDQASLHAFADSLERDLDAPSPTHEGDQRRASSRKRSRSDRAVTGPSAQQSFEVRAPEQRDALHLPLSWRVKRPRTSIGGGLPDPGTPTAADLAASSTVMREQDEDPFAGAADDFPAFNEEELAWLMELLPQDRGRAPPTDVSLGDELHLDGEDVAMAHADALDDFDLDMLGDGDSPGPGFTPHDSAPYGALDMADFEFEQMFTDALGIDEYGGGRDYKDDDDK
ATGGTGGATCTACGCACGCTCGGCTACAGCCAGCAGCAACAGGAGAAGATCAAACCGAAGGTTCGTTCGACAGTGGCGCAGCACCACGAGGCACTGGTCGGCCATGGGTTTACACACGCGCACATCGTTGCGCTCAGCCAACACCCGGCAGCGTTAGGGACCGTCGCTGTCAAGTATCAGGACATGATCGCAGCGTTGCCAGAGGCGACACACGAAGCGATCGTTGGCGTCGGCAAACAGTGGTCCGGCGCACGCGCCCTGGAGGCCTTGCTCACGGTGGCGGGAGAGTTGAGAGGTCCACCGTTACAGTTGGACACAGGCCAACTTCTCAAGATTGCAAAACGTGGCGGCGTGACCGCAGTGGAGGCAGTGCATGCATGGCGCAATGCACTGACGGGGGCCCCCCTGAACCTTACTCCCGAACAAGTAGTAGCGATAGCCAGTAATAACGGAGGTAAACAAGCCTTGGAGACGGTCCAAAGGTTGCTCCCGGTCTTGTGTCAGGCACATGGGCTGACGCCTCAACAGGTCGTCGCGATAGCGTCTAATAATGGAGGAAAGCAAGCTCTGGAAACCGTCCAGCGACTCCTTCCGGTTCTGTGCCAGGCTCATGGTCTGACTCCGCAGCAAGTCGTTGCTATAGCGTCCAACATCGGAGGCAAACAGGCCCTGGAGACCGTGCAGCGGTTGTTGCCTGTGCTTTGCCAAGCCCACGGGCTTACGCCTGAGCAAGTGGTGGCGATTGCCAGTAACAACGGCGGCAAACAAGCCCTTGAGACTGTGCAGAGGCTCTTGCCGGTACTCTGCCAAGCACACGGCTTGACCCCCGAGCAGGTTGTAGCCATAGCTAGTCACGACGGGGGTAAGCAAGCGTTGGAAACGGTGCAAGCACTTCTCCCCGTTCTCTGTCAAGCGCATGGACTTACCCCGGAACAGGTGGTCGCCATTGCAAGCCATGATGGAGGAAAGCAGGCGCTCGAAACAGTCCAGGCACTTTTGCCCGTACTTTGTCAAGCTCACGGTCTCACCCCGGAACAGGTGGTAGCCATTGCATCTAACATCGGAGGTAAGCAAGCATTGGAAACGGTTCAGGCCCTGTTGCCTGTACTTTGCCAGGCGCACGGTCTGACACCTGAGCAGGTTGTCGCCATCGCTAGCAACGGAGGTGGGAAACAGGCACTTGAAACTGTGCAGAGGCTTCTGCCGGTGCTGTGCCAAGCGCATGGCCTTACACCCGAGCAAGTAGTGGCTATTGCGAGTCATGATGGAGGCAAGCAAGCGCTGGAGACTGTCCAACGACTTCTTCCGGTCTTGTGTCAGGCACATGGATTGACCCCTCAACAAGTCGTGGCGATAGCTAGCAACGGCGGTGGAAAACAGGCCCTCGAAACCGTCCAGCGACTGCTCCCCGTACTGTGTCAAGCCCATGGACTTACCCCAGAACAAGTTGTGGCGATTGCCTCTAACAATGGTGGGAAGCAAGCTCTTGAGACGGTGCAGGCGTTGTTGCCCGTGCTTTGTCAAGCTCACGGGCTCACGCCAGAGCAAGTGGTCGCTATCGCGAGTAATAAAGGGGGCAAACAAGCCTTGGAGACAGTGCAAAGGCTCCTGCCAGTGCTCTGCCAGGCTCATGGTTTGACACCCGAACAGGTAGTTGCAATAGCGAGTCATGATGGCGGAAAGCAAGCTCTTGAAACTGTGCAGCGGCTGTTGCCTGTACTGTGTCAAGCCCACGGGCTGACACCGGAACAAGTTGTAGCGATCGCTAGCCACGATGGCGGGAAACAAGCTCTGGAAACGGTACAGAGACTCCTCCCAGTGCTTTGTCAGGCACACGGCCTCACGCCAGAGCAGGTTGTCGCCATCGCGTCAAACAATGGTGGAAAGCAGGCCCTGGAGACAGTCCAACGGTTGCTGCCGGTCCTTTGCCAGGCTCACGGGTTGACCCCCCAGCAGGTCGTGGCCATTGCCTCAAACAAGGGCGGTAGGCCAGCATTGGAGACGGTGCAGAGGCTTCTGCCTGTGCTCTGCCAAGCGCATGGACTCACCCCCGAGCAAGTGGTTGCTATCGCAAGTAACAACGGAGGGAAACAAGCGCTCGAAACCGTGCAAAGGTTGCTCCCCGTTCTCTGTCAGGCGCACGGTCTTACGCCACAACAGGTGGTGGCGATTGCATCTAATGGAGGCGGACGCCCTGCCTTGGAGAGCATTGTGGCCCAGCTGTCCAGGCCGGACCCTGCCCTGGCCGCGTTAACCAACGACCACCTCGTCGCCTTGGCCTGCCTCGGCGGACGTCCTGCGCTGGATGCAGTGAAAAAGGGATTGCCGCACGCGCCGGCCTTGATCAAAAGAACCAATCGCCGTATTCCCGAACGCACATCCCATCGCGTTGCCGACCACGCGCAAGTGGTTCGCGTGCTGGGTTTTTTCCAGTGCCACTCCCACCCAGCGCAAGCATTTGATGACGCCATGACGCAGTTCGGGATGAGCAGGCACGGGTTGTTACAGCTCTTTCGCAGAGTGGGCGTCACCGAACTCGAAGCCCGCAGTGGAACGCTCCCCCCAGCCTCGCAGCGTTGGGACCGTATCCTCCAGGCATCAGGGATGAAAAGGGCCAAACCGTCCCCTACTTCAACTCAAACGCCGGACCAGGCGTCTTTGCATGCATTCGCCGATTCGCTGGAGCGTGACCTTGATGCGCCCAGCCCAACGCACGAGGGAGATCAGAGGCGGGCAAGCAGCCGTAAACGGTCCCGATCGGATCGTGCTGTCACCGGTCCCTCCGCACAGCAATCGTTCGAGGTGCGCGCTCCCGAACAGCGCGATGCGCTGCATTTGCCCCTCAGTTGGAGGGTAAAACGCCCGCGTACCAGTATCGGGGGCGGCCTCCCGGATCCTGGTACGCCCACGGCTGCCGACCTGGCAGCGTCCAGCACCGTGATGCGGGAACAAGATGAGGACCCCTTCGCAGGGGCAGCGGATGATTTCCCGGCATTCAACGAAGAGGAGCTCGCATGGTTGATGGAGCTATTGCCTCAGGACCGCGGCCGCGCCCCCCCGACCGATGTCAGCCTGGGGGACGAGCTCCACTTAGACGGCGAGGACGTGGCGATGGCGCATGCCGACGCGCTAGACGATTTCGATCTGGACATGTTGGGGGACGGGGATTCCCCGGGTCCGGGATTTACCCCCCACGACTCCGCCCCCTACGGCGCTCTGGATATGGCCGACTTCGAGTTTGAGCAGATGTTTACCGATGCCCTTGGAATTGACGAGTACGGTGGCGGCCGCGACTACAAGGACGACGATGACAAG
MAPKKKRKVGIDGVPAAVDLRTLGYSQQQQEKIKPKVRSTVAQHHEALVGHGFTHAHIVALSQHPAALGTVAVKYQDMIAALPEATHEAIVGVGKQWSGARALEALLTVAGELRGPPLQLDTGQLLKIAKRGGVTAVEAVHAWRNALTGAPLNLTPEQVVAIASNNGGKQALETVQRLLPVLCQAHGLTPQQVVAIASNNGGKQALETVQRLLPVLCQAHGLTPQQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNNGGKQALETVQRLLPVLCQAHGLTPEQVVAIASHDGGKQALETVQALLPVLCQAHGLTPEQVVAIASHDGGKQALETVQALLPVLCQAHGLTPEQVVAIASNIGGKQALETVQALLPVLCQAHGLTPEQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPQQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNNGGKQALETVQALLPVLCQAHGLTPEQVVAIASNKGGKQALETVQRLLPVLCQAHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNNGGKQALETVQRLLPVLCQAHGLTPQQVVAIASNKGGRPALETVQRLLPVLCQAHGLTPEQVVAIASNNGGKQALETVQRLLPVLCQAHGLTPQQVVAIASNGGGRPALESIVAQLSRPDPALAALTNDHLVALACLGGRPALDAVKKGLPHAPALIKRTNRRIPERTSHRVADHAQVVRVLGFFQCHSHPAQAFDDAMTQFGMSGSRGRAPPTDVSLGDELHLDGEDVAMAHADALDDFDLDMLGDGDSPGPGFTPHDSAPYGALDMADFEFEQMFTDALGIDEYGGGRDYKDDDDK
ATGGCCCCCAAGAAGAAGAGGAAGGTGGGAATCGATGGGGTACCCGCCGCTGTGGATCTACGCACGCTCGGCTACAGCCAGCAGCAACAGGAGAAGATCAAACCGAAGGTTCGTTCGACAGTGGCGCAGCACCACGAGGCACTGGTCGGCCATGGGTTTACACACGCGCACATCGTTGCGCTCAGCCAACACCCGGCAGCGTTAGGGACCGTCGCTGTCAAGTATCAGGACATGATCGCAGCGTTGCCAGAGGCGACACACGAAGCGATCGTTGGCGTCGGCAAACAGTGGTCCGGCGCACGCGCCCTGGAGGCCTTGCTCACGGTGGCGGGAGAGTTGAGAGGTCCACCGTTACAGTTGGACACAGGCCAACTTCTCAAGATTGCAAAACGTGGCGGCGTGACCGCAGTGGAGGCAGTGCATGCATGGCGCAATGCACTGACGGGGGCCCCCCTGAACCTTACTCCCGAACAAGTAGTAGCGATAGCCAGTAATAACGGAGGTAAACAAGCCTTGGAGACGGTCCAAAGGTTGCTCCCGGTCTTGTGTCAGGCACATGGGCTGACGCCTCAACAGGTCGTCGCGATAGCGTCTAATAATGGAGGAAAGCAAGCTCTGGAAACCGTCCAGCGACTCCTTCCGGTTCTGTGCCAGGCTCATGGTCTGACTCCGCAGCAAGTCGTTGCTATAGCGTCCAACATCGGAGGCAAACAGGCCCTGGAGACCGTGCAGCGGTTGTTGCCTGTGCTTTGCCAAGCCCACGGGCTTACGCCTGAGCAAGTGGTGGCGATTGCCAGTAACAACGGCGGCAAACAAGCCCTTGAGACTGTGCAGAGGCTCTTGCCGGTACTCTGCCAAGCACACGGCTTGACCCCCGAGCAGGTTGTAGCCATAGCTAGTCACGACGGGGGTAAGCAAGCGTTGGAAACGGTGCAAGCACTTCTCCCCGTTCTCTGTCAAGCGCATGGACTTACCCCGGAACAGGTGGTCGCCATTGCAAGCCATGATGGAGGAAAGCAGGCGCTCGAAACAGTCCAGGCACTTTTGCCCGTACTTTGTCAAGCTCACGGTCTCACCCCGGAACAGGTGGTAGCCATTGCATCTAACATCGGAGGTAAGCAAGCATTGGAAACGGTTCAGGCCCTGTTGCCTGTACTTTGCCAGGCGCACGGTCTGACACCTGAGCAGGTTGTCGCCATCGCTAGCAACGGAGGTGGGAAACAGGCACTTGAAACTGTGCAGAGGCTTCTGCCGGTGCTGTGCCAAGCGCATGGCCTTACACCCGAGCAAGTAGTGGCTATTGCGAGTCATGATGGAGGCAAGCAAGCGCTGGAGACTGTCCAACGACTTCTTCCGGTCTTGTGTCAGGCACATGGATTGACCCCTCAACAAGTCGTGGCGATAGCTAGCAACGGCGGTGGAAAACAGGCCCTCGAAACCGTCCAGCGACTGCTCCCCGTACTGTGTCAAGCCCATGGACTTACCCCAGAACAAGTTGTGGCGATTGCCTCTAACAATGGTGGGAAGCAAGCTCTTGAGACGGTGCAGGCGTTGTTGCCCGTGCTTTGTCAAGCTCACGGGCTCACGCCAGAGCAAGTGGTCGCTATCGCGAGTAATAAAGGGGGCAAACAAGCCTTGGAGACAGTGCAAAGGCTCCTGCCAGTGCTCTGCCAGGCTCATGGTTTGACACCCGAACAGGTAGTTGCAATAGCGAGTCATGATGGCGGAAAGCAAGCTCTTGAAACTGTGCAGCGGCTGTTGCCTGTACTGTGTCAAGCCCACGGGCTGACACCGGAACAAGTTGTAGCGATCGCTAGCCACGATGGCGGGAAACAAGCTCTGGAAACGGTACAGAGACTCCTCCCAGTGCTTTGTCAGGCACACGGCCTCACGCCAGAGCAGGTTGTCGCCATCGCGTCAAACAATGGTGGAAAGCAGGCCCTGGAGACAGTCCAACGGTTGCTGCCGGTCCTTTGCCAGGCTCACGGGTTGACCCCCCAGCAGGTCGTGGCCATTGCCTCAAACAAGGGCGGTAGGCCAGCATTGGAGACGGTGCAGAGGCTTCTGCCTGTGCTCTGCCAAGCGCATGGACTCACCCCCGAGCAAGTGGTTGCTATCGCAAGTAACAACGGAGGGAAACAAGCGCTCGAAACCGTGCAAAGGTTGCTCCCCGTTCTCTGTCAGGCGCACGGTCTTACGCCACAACAGGTGGTGGCGATTGCATCTAATGGAGGCGGACGCCCTGCCTTGGAGAGCATTGTGGCCCAGCTGTCCAGGCCGGACCCTGCCCTGGCCGCGTTAACCAACGACCACCTCGTCGCCTTGGCCTGCCTCGGCGGACGTCCTGCGCTGGATGCAGTGAAAAAGGGATTGCCGCACGCGCCGGCCTTGATCAAAAGAACCAATCGCCGTATTCCCGAACGCACATCCCATCGCGTTGCCGACCACGCGCAAGTGGTTCGCGTGCTGGGTTTTTTCCAGTGCCACTCCCACCCAGCGCAAGCATTTGATGACGCCATGACGCAGTTCGGGATGAGCGGATCCCGCGGCCGCGCCCCCCCGACCGATGTCAGCCTGGGGGACGAGCTCCACTTAGACGGCGAGGACGTGGCGATGGCGCATGCCGACGCGCTAGACGATTTCGATCTGGACATGTTGGGGGACGGGGATTCCCCGGGTCCGGGATTTACCCCCCACGACTCCGCCCCCTACGGCGCTCTGGATATGGCCGACTTCGAGTTTGAGCAGATGTTTACCGATGCCCTTGGAATTGACGAGTACGGTGGCGGCCGCGACTACAAGGACGACGATGACAAG
MAPKKKRKVGIDGVPAAVDLRTLGYSQQQQEKIKPKVRSTVAQHHEALVGHGFTHAHIVALSQHPAALGTVAVKYQDMIAALPEATHEAIVGVGKQWSGARALEALLTVAGELRGPPLQLDTGQLLKIAKRGGVTAVEAVHAWRNALTGAPLNLTPEQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNIGGKQALETVQRLLPVLRQAHGLTPEQVVAIASNGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNGGGRPALESIVAQLSRPDPALAALTNDHLVALACLGGRPALDAVKKGLPHAPALIKRTNRRIPERTSHRVADHAQVVRVLGFFQCHSHPAQAFDDAMTQFGMSRHGLLQLFRRVGVTELEARSGTLPPASQRWDRILQASGMKRAKPSPTSTQTPDQASLHAFADSLERDLDAPSPTHEGDQRRASSRKRSRSDRAVTGPSAQQSFEVRAPEQRDALHLPLSWRVKRPRTSIGGGLPDPGTPTAADLAASSTVMREQDEDPFAGAADDFPAFNEEELAWLMELLPQDRGRAPPTDVSLGDELHLDGEDVAMAHADALDDFDLDMLGDGDSPGPGFTPHDSAPYGALDMADFEFEQMFTDALGIDEYGGGRDYKDDDDK
ATGGCCCCCAAGAAGAAGAGGAAGGTGGGAATCGATGGGGTACCCGCCGCTGTGGATCTACGCACGCTCGGCTACAGCCAGCAGCAACAGGAGAAGATCAAACCGAAGGTTCGTTCGACAGTGGCGCAGCACCACGAGGCACTGGTCGGCCATGGGTTTACACACGCGCACATCGTTGCGCTCAGCCAACACCCGGCAGCGTTAGGGACCGTCGCTGTCAAGTATCAGGACATGATCGCAGCGTTGCCAGAGGCGACACACGAAGCGATCGTTGGCGTCGGCAAACAGTGGTCCGGCGCACGCGCCCTGGAGGCCTTGCTCACGGTGGCGGGAGAGTTGAGAGGTCCACCGTTACAGTTGGACACAGGCCAACTTCTCAAGATTGCAAAACGTGGCGGCGTGACCGCAGTGGAGGCAGTGCATGCATGGCGCAATGCACTGACGGGGGCCCCCCTGAACCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCAATATTGGTGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCAATGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCAATATTGGTGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCAATATTGGTGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGCACAGGTGGTGGCCATCGCCAGCAATATTGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGCGCCAGGCCCATGGCCTGACCCCGGAGCAGGTCGTGGCCATCGCCAGCAATGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGCACAGGTGGTGGCCATCGCCAGCAATATTGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTCGTGGCCATCGCCAGCCACGATGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCCACGATGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGACCAGGTGGTGGCCATCGCCAGCAATGGCGGTGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGACCAGGTGGTGGCCATCGCCAGCAATGGCGGTGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCCACGATGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCAATGGCGGCGGCAGGCCGGCGCTGGAGAGCATTGTTGCCCAGTTATCTCGCCCTGATCCGGCGTTGGCCGCGTTAACCAACGACCACCTCGTCGCCTTGGCCTGCCTCGGCGGACGTCCTGCGCTGGATGCAGTGAAAAAGGGATTGCCGCACGCGCCGGCCTTGATCAAAAGAACCAATCGCCGTATTCCCGAACGCACATCCCATCGCGTTGCCGACCACGCGCAAGTGGTTCGCGTGCTGGGTTTTTTCCAGTGCCACTCCCACCCAGCGCAAGCATTTGATGACGCCATGACGCAGTTCGGGATGAGCAGGCACGGGTTGTTACAGCTCTTTCGCAGAGTGGGCGTCACCGAACTCGAAGCCCGCAGTGGAACGCTCCCCCCAGCCTCGCAGCGTTGGGACCGTATCCTCCAGGCATCAGGGATGAAAAGGGCCAAACCGTCCCCTACTTCAACTCAAACGCCGGACCAGGCGTCTTTGCATGCATTCGCCGATTCGCTGGAGCGTGACCTTGATGCGCCCAGCCCAACGCACGAGGGAGATCAGAGGCGGGCAAGCAGCCGTAAACGGTCCCGATCGGATCGTGCTGTCACCGGTCCCTCCGCACAGCAATCGTTCGAGGTGCGCGCTCCCGAACAGCGCGATGCGCTGCATTTGCCCCTCAGTTGGAGGGTAAAACGCCCGCGTACCAGTATCGGGGGCGGCCTCCCGGATCCTGGTACGCCCACGGCTGCCGACCTGGCAGCGTCCAGCACCGTGATGCGGGAACAAGATGAGGACCCCTTCGCAGGGGCAGCGGATGATTTCCCGGCATTCAACGAAGAGGAGCTCGCATGGTTGATGGAGCTATTGCCTCAGGACCGCGGCCGCGCCCCCCCGACCGATGTCAGCCTGGGGGACGAGCTCCACTTAGACGGCGAGGACGTGGCGATGGCGCATGCCGACGCGCTAGACGATTTCGATCTGGACATGTTGGGGGACGGGGATTCCCCGGGTCCGGGATTTACCCCCCACGACTCCGCCCCCTACGGCGCTCTGGATATGGCCGACTTCGAGTTTGAGCAGATGTTTACCGATGCCCTTGGAATTGACGAGTACGGTGGCGGCCGCGACTACAAGGACGACGATGACAAG
MVDLRTLGYSQQQQEKIKPKVRSTVAQHHEALVGHGFTHAHIVALSQHPAALGTVAVKYQDMIAALPEATHEAIVGVGKQWSGARALEALLTVAGELRGPPLQLDTGQLLKIAKRGGVTAVEAVHAWRNALTGAPLNLTPEQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNIGGKQALETVQRLLPVLRQAHGLTPEQVVAIASNGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNGGGRPALESIVAQLSRPDPALAALTNDHLVALACLGGRPALDAVKKGLPHAPALIKRTNRRIPERTSHRVADHAQVVRVLGFFQCHSHPAQAFDDAMTQFGMSRHGLLQLFRRVGVTELEARSGTLPPASQRWDRILQASGGSGHRGRAPPTDVSLGDELHLDGEDVAMAHADALDDFDLDMLGDGDSPGPGFTPHDSAPYGALDMADFEFEQMFTDALGIDEYGGGRDYKDDDDK
ATGGTGGATCTACGCACGCTCGGCTACAGCCAGCAGCAACAGGAGAAGATCAAACCGAAGGTTCGTTCGACAGTGGCGCAGCACCACGAGGCACTGGTCGGCCATGGGTTTACACACGCGCACATCGTTGCGCTCAGCCAACACCCGGCAGCGTTAGGGACCGTCGCTGTCAAGTATCAGGACATGATCGCAGCGTTGCCAGAGGCGACACACGAAGCGATCGTTGGCGTCGGCAAACAGTGGTCCGGCGCACGCGCCCTGGAGGCCTTGCTCACGGTGGCGGGAGAGTTGAGAGGTCCACCGTTACAGTTGGACACAGGCCAACTTCTCAAGATTGCAAAACGTGGCGGCGTGACCGCAGTGGAGGCAGTGCATGCATGGCGCAATGCACTGACGGGTGCCCCCCTGAACCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCAATATTGGTGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCAATGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCAATATTGGTGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCAATATTGGTGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGCACAGGTGGTGGCCATCGCCAGCAATATTGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGCGCCAGGCCCATGGCCTGACCCCGGAGCAGGTCGTGGCCATCGCCAGCAATGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGCACAGGTGGTGGCCATCGCCAGCAATATTGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTCGTGGCCATCGCCAGCCACGATGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCCACGATGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGACCAGGTGGTGGCCATCGCCAGCAATGGCGGTGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGACCAGGTGGTGGCCATCGCCAGCAATGGCGGTGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCCACGATGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCAATGGCGGCGGCAGGCCGGCGCTGGAGAGCATTGTTGCCCAGTTATCTCGCCCTGATCCGGCGTTGGCCGCGTTGACCAACGACCACCTCGTCGCCTTGGCCTGCCTCGGCGGACGTCCTGCGCTGGATGCAGTGAAAAAGGGATTGCCGCACGCGCCGGCCTTGATCAAAAGAACCAATCGCCGTATTCCCGAACGCACATCCCATCGCGTTGCCGACCACGCGCAAGTGGTTCGCGTGCTGGGTTTTTTCCAGTGCCACTCCCACCCAGCGCAAGCATTTGATGACGCCATGACGCAGTTCGGGATGAGCAGGCACGGGTTGTTACAGCTCTTTCGCAGAGTGGGCGTCACCGAACTCGAAGCCCGCAGTGGAACGCTCCCCCCAGCCTCGCAGCGTTGGGACCGTATCCTCCAGGCATCGGGGGGATCCGGCCACCGCGGCCGCGCCCCCCCGACCGATGTCAGCCTGGGGGACGAGCTCCACTTAGACGGCGAGGACGTGGCGATGGCGCATGCCGACGCGCTAGACGATTTCGATCTGGACATGTTGGGGGACGGGGATTCCCCGGGTCCGGGATTTACCCCCCACGACTCCGCCCCCTACGGCGCTCTGGATATGGCCGACTTCGAGTTTGAGCAGATGTTTACCGATGCCCTTGGAATTGACGAGTACGGTGGCGGCCGCGACTACAAGGACGACGATGACAAG
MVDLRTLGYSQQQQEKIKPKVRSTVAQHHEALVGHGFTHAHIVALSQHPAALGTVAVKYQDMIAALPEATHEAIVGVGKQWSGARALEALLTVAGELRGPPLQLDTGQLLKIAKRGGVTAVEAVHAWRNALTGAPLNLTPEQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNIGGKQALETVQRLLPVLRQAHGLTPEQVVAIASNGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNGGGRPALESIVAQLSRPDPALAALTNDHLVALACLGGRPALDAVKKGLPHAPALIKRTNRRIPERTSHRVADHAQVVRVLGFFQCHSHPAQAFDDAMTQFGMSGSRGRAPPTDVSLGDELHLDGEDVAMAHADALDDFDLDMLGDGDSPGPGFTPHDSAPYGALDMADFEFEQMFTDALGIDEYGGGRDYKDDDDK
ATGGTGGATCTACGCACGCTCGGCTACAGCCAGCAGCAACAGGAGAAGATCAAACCGAAGGTTCGTTCGACAGTGGCGCAGCACCACGAGGCACTGGTCGGCCATGGGTTTACACACGCGCACATCGTTGCGCTCAGCCAACACCCGGCAGCGTTAGGGACCGTCGCTGTCAAGTATCAGGACATGATCGCAGCGTTGCCAGAGGCGACACACGAAGCGATCGTTGGCGTCGGCAAACAGTGGTCCGGCGCACGCGCCCTGGAGGCCTTGCTCACGGTGGCGGGAGAGTTGAGAGGTCCACCGTTACAGTTGGACACAGGCCAACTTCTCAAGATTGCAAAACGTGGCGGCGTGACCGCAGTGGAGGCAGTGCATGCATGGCGCAATGCACTGACGGGGGCCCCCCTGAACCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCAATATTGGTGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCAATGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCAATATTGGTGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCAATATTGGTGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGCACAGGTGGTGGCCATCGCCAGCAATATTGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGCGCCAGGCCCATGGCCTGACCCCGGAGCAGGTCGTGGCCATCGCCAGCAATGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGCACAGGTGGTGGCCATCGCCAGCAATATTGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTCGTGGCCATCGCCAGCCACGATGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCCACGATGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGACCAGGTGGTGGCCATCGCCAGCAATGGCGGTGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGACCAGGTGGTGGCCATCGCCAGCAATGGCGGTGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCCACGATGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCAATGGCGGCGGCAGGCCGGCGCTGGAGAGCATTGTTGCCCAGTTATCTCGCCCTGATCCGGCGTTGGCCGCGTTAACCAACGACCACCTCGTCGCCTTGGCCTGCCTCGGCGGACGTCCTGCGCTGGATGCAGTGAAAAAGGGATTGCCGCACGCGCCGGCCTTGATCAAAAGAACCAATCGCCGTATTCCCGAACGCACATCCCATCGCGTTGCCGACCACGCGCAAGTGGTTCGCGTGCTGGGTTTTTTCCAGTGCCACTCCCACCCAGCGCAAGCATTTGATGACGCCATGACGCAGTTCGGGATGAGCGGATCCCGCGGCCGCGCCCCCCCGACCGATGTCAGCCTGGGGGACGAGCTCCACTTAGACGGCGAGGACGTGGCGATGGCGCATGCCGACGCGCTAGACGATTTCGATCTGGACATGTTGGGGGACGGGGATTCCCCGGGTCCGGGATTTACCCCCCACGACTCCGCCCCCTACGGCGCTCTGGATATGGCCGACTTCGAGTTTGAGCAGATGTTTACCGATGCCCTTGGAATTGACGAGTACGGTGGCGGCCGCGACTACAAGGACGACGATGACAAG
MVDLRTLGYSQQQQEKIKPKVRSTVAQHHEALVGHGFTHAHIVALSQHPAALGTVAVKYQDMIAALPEATHEAIVGVGKQWSGARALEALLTVAGELRGPPLQLDTGQLLKIAKRGGVTAVEAVHAWRNALTGAPLNLTPEQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNIGGKQALETVQRLLPVLRQAHGLTPEQVVAIASNGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNGGGRPALESIVAQLSRPDPALAALTNDHLVAGSRGRAPPTDVSLGDELHLDGEDVAMAHADALDDFDLDMLGDGDSPGPGFTPHDSAPYGALDMADFEFEQMFTDALGIDEYGGGRDYKDDDDK
ATGGTGGATCTACGCACGCTCGGCTACAGCCAGCAGCAACAGGAGAAGATCAAACCGAAGGTTCGTTCGACAGTGGCGCAGCACCACGAGGCACTGGTCGGCCATGGGTTTACACACGCGCACATCGTTGCGCTCAGCCAACACCCGGCAGCGTTAGGGACCGTCGCTGTCAAGTATCAGGACATGATCGCAGCGTTGCCAGAGGCGACACACGAAGCGATCGTTGGCGTCGGCAAACAGTGGTCCGGCGCACGCGCCCTGGAGGCCTTGCTCACGGTGGCGGGAGAGTTGAGAGGTCCACCGTTACAGTTGGACACAGGCCAACTTCTCAAGATTGCAAAACGTGGCGGCGTGACCGCAGTGGAGGCAGTGCATGCATGGCGCAATGCACTGACGGGGGCCCCCCTGAACCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCAATATTGGTGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCAATGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCAATATTGGTGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCAATATTGGTGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGCACAGGTGGTGGCCATCGCCAGCAATATTGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGCGCCAGGCCCATGGCCTGACCCCGGAGCAGGTCGTGGCCATCGCCAGCAATGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGCACAGGTGGTGGCCATCGCCAGCAATATTGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTCGTGGCCATCGCCAGCCACGATGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCCACGATGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGACCAGGTGGTGGCCATCGCCAGCAATGGCGGTGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGACCAGGTGGTGGCCATCGCCAGCAATGGCGGTGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCCACGATGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCAATGGCGGCGGCAGGCCGGCGCTGGAGAGCATTGTTGCCCAGTTATCTCGCCCTGATCCGGCGTTGGCCGCGTTAACCAACGACCACCTCGTCGCCGGATCCCGCGGCCGCGCCCCCCCGACCGATGTCAGCCTGGGGGACGAGCTCCACTTAGACGGCGAGGACGTGGCGATGGCGCATGCCGACGCGCTAGACGATTTCGATCTGGACATGTTGGGGGACGGGGATTCCCCGGGTCCGGGATTTACCCCCCACGACTCCGCCCCCTACGGCGCTCTGGATATGGCCGACTTCGAGTTTGAGCAGATGTTTACCGATGCCCTTGGAATTGACGAGTACGGTGGCGGCCGCGACTACAAGGACGACGATGACAAG
MVDLRTLGYSQQQQEKIKPKVRSTVAQHHEALVGHGFTHAHIVALSQHPAALGTVAVKYQDMIAALPEATHEAIVGVGKQWSGARALEALLTVAGELRGPPLQLDTGQLLKIAKRGGVTAVEAVHAWRNALTGAPLNALTNDHLVALACLGGRPALDAVKKGLPHAPALIKRTNRRIPERTSHRVADHAQVVRVLGFFQCHSHPAQAFDDAMTQFGMSRHGLLQLFRRVGVTELEARSGTLPPASQRWDRILQASGMKRAKPSPTSTQTPDQASLHAFADSLERDLDAPSPTHEGDQRRASSRKRSRSDRAVTGPSAQQSFEVRAPEQRDALHLPLSWRVKRPRTSIGGGLPDPGTPTAADLAASSTVMREQDEDPFAGAADDFPAFNEEELAWLMELLPQDRGRAPPTDVSLGDELHLDGEDVAMAHADALDDFDLDMLGDGDSPGPGFTPHDSAPYGALDMADFEFEQMFTDALGIDEYGGGRDYKDDDDK
ATGGTGGATCTACGCACGCTCGGCTACAGCCAGCAGCAACAGGAGAAGATCAAACCGAAGGTTCGTTCGACAGTGGCGCAGCACCACGAGGCACTGGTCGGCCATGGGTTTACACACGCGCACATCGTTGCGCTCAGCCAACACCCGGCAGCGTTAGGGACCGTCGCTGTCAAGTATCAGGACATGATCGCAGCGTTGCCAGAGGCGACACACGAAGCGATCGTTGGCGTCGGCAAACAGTGGTCCGGCGCACGCGCCCTGGAGGCCTTGCTCACGGTGGCGGGAGAGTTGAGAGGTCCACCGTTACAGTTGGACACAGGCCAACTTCTCAAGATTGCAAAACGTGGCGGCGTGACCGCAGTGGAGGCAGTGCATGCATGGCGCAATGCACTGACGGGGGCCCCCCTGAACGCGTTAACCAACGACCACCTCGTCGCCTTGGCCTGCCTCGGCGGACGTCCTGCGCTGGATGCAGTGAAAAAGGGATTGCCGCACGCGCCGGCCTTGATCAAAAGAACCAATCGCCGTATTCCCGAACGCACATCCCATCGCGTTGCCGACCACGCGCAAGTGGTTCGCGTGCTGGGTTTTTTCCAGTGCCACTCCCACCCAGCGCAAGCATTTGATGACGCCATGACGCAGTTCGGGATGAGCAGGCACGGGTTGTTACAGCTCTTTCGCAGAGTGGGCGTCACCGAACTCGAAGCCCGCAGTGGAACGCTCCCCCCAGCCTCGCAGCGTTGGGACCGTATCCTCCAGGCATCAGGGATGAAAAGGGCCAAACCGTCCCCTACTTCAACTCAAACGCCGGACCAGGCGTCTTTGCATGCATTCGCCGATTCGCTGGAGCGTGACCTTGATGCGCCCAGCCCAACGCACGAGGGAGATCAGAGGCGGGCAAGCAGCCGTAAACGGTCCCGATCGGATCGTGCTGTCACCGGTCCCTCCGCACAGCAATCGTTCGAGGTGCGCGCTCCCGAACAGCGCGATGCGCTGCATTTGCCCCTCAGTTGGAGGGTAAAACGCCCGCGTACCAGTATCGGGGGCGGCCTCCCGGATCCTGGTACGCCCACGGCTGCCGACCTGGCAGCGTCCAGCACCGTGATGCGGGAACAAGATGAGGACCCCTTCGCAGGGGCAGCGGATGATTTCCCGGCATTCAACGAAGAGGAGCTCGCATGGTTGATGGAGCTATTGCCTCAGGACCGCGGCCGCGCCCCCCCGACCGATGTCAGCCTGGGGGACGAGCTCCACTTAGACGGCGAGGACGTGGCGATGGCGCATGCCGACGCGCTAGACGATTTCGATCTGGACATGTTGGGGGACGGGGATTCCCCGGGTCCGGGATTTACCCCCCACGACTCCGCCCCCTACGGCGCTCTGGATATGGCCGACTTCGAGTTTGAGCAGATGTTTACCGATGCCCTTGGAATTGACGAGTACGGTGGCGGCCGCGACTACAAGGACGACGATGACAAG
図37に説明される実験のために使用したドナー(配列番号318)
AGCGCCCAATACGCAAACCGCCTCTCCCCGCGCGTTGGCCGATTCATTAATGCAGCTGGCACGACAGGTTTCCCGACTGGAAAGCGGGCAGTGAGCGCAACGCAATTAATGTGAGTTAGCTCACTCATTAGGCACCCCAGGCTTTACACTTTATGCTTCCGGCTCGTATGTTGTGTGGAATTGTGAGCGGATAACAATTTCACACAGGAAACAGCTATGACCATGATTACGCCAAGCTCAGAATTAACCCTCACTAAAGGGACTAGTCCTGCAGGTTTAAACGAATTCGCCCTTGATACTTATTAACCATACCTTGGAGGGGAAATCACACATGAAAAGTGTCATTTCTTTACTAATCATATTCATGTCTTTTCTCCCCATAGCAAGACAAAGACCTGTTTTAAACACATTTACAACCTATATGTTGCCTTGTACTAGGTAAAAAGTTGTACATTTCTGAAATAATTTTGGTATTTCTGTTCAGATCACTAAACTCAAGAATCAGCAATTCTCTGAGGCTTTCTTTTAAATATACATAAGGAACTTTCGGAGTGAAGGGAGAGTTTGTCAATAACTTGATGCATGTGAAGGGGAGATAAAAAGGTTGCTATTTTTCATCAACATATTTTGATTTGGCTTTCTATAATTGATGGGCTTAAAAGATCTAATCTACTTTAAACAGATGCCAAATAAATGGATGAATCTTAGACCCTCTATAACAGTAACTTCCTTTTAAAAAAGACCTCTCCCACCCCACCCCCAGCCCAGGCTGTGTATGAAAACTAAGCCATGTGCACAACTCTGACTGGGTCACCAGCCCACTTGAGTCCGTGTCACAAGCCCACAGATATTTCCTGCTCCCCAGTGGATCGGGTGTAAACTGAGCTTGCTCGCTCGGGAGCCTCTTGCTGGAAAATAGAACAGCATTTGCAGAAGCGTTTGGCAATGTGCTTTTGGAAGAAGACTAAGAGGTAGTTTCTGAACTTCTCCCCGACAAAGGCATAGATGATGGGGTTGATGCAGCAGTGCGTCATCCCAAGAGTCTCTGTCACCTGCATAGCTTGGTCCAACCTGTTAGAGCTACTGCAATTATTCAGGCCAAAGAATTCCTGGAAGGTGTTCAGGAGAAGGACAATGTTGTAGGGAGCCCAGAAGAGAAAATAAACAATCATGATGGTGAAGATAAGCCTCACAGCCCTGTGCCTCTTCTTCTCATTTCGACACCGAAGCAGAGTTTTTAGGATTCCCGAGTAGCAGATGACCATGACAAGCAGCGGCAGGACCAGCCCCAAGATGACTATCTTTAATGTCTGGAAATTCTTCCAGAATTGATACTGACTGTATGGAAAATGAGAGCTGCAGGTGTAATGAAGACCTTCTTTTTGAGATCTGGTAAAGATGATTCCTGGGAGAGACGCAAACACAGCCACCACCCAAGTGATCACACTTGTCACCACCCCAAAGGTGACCGTCCTGGCTTTTAAAGCAAACACAGCATGGACGACAGCCAGGTACCTATCGATTGTCAGGAGGATGATGAAGAAGATTCCAGAGAAGAAGCCTATAAAATAGAGCCCTGTCAAGAGTTGACACATTGTATTTCCAAAGTCCCACTGGGCGGCAGCATAGTGAGCCCAGAAGGGGACAGTAAGAAGGAAAAACAGGTCAGAGATGGCCAGGTTGAGCAGGTAGATGTCAGTCATGCTCTTCAGCCTTTTGCAGTTTTCTAGACGAGGCATCCAGTCCAGACGCCATCAGGGCATACTCACTGATCTAGATGAGGATGACCAGCATGTTGCCCACAAAACCAAAGATGAACACCAGTGAGTAGAGCGGAGGCAGGAGGCGGGCTGCGATTTGCTTCACATTGATTTTTTGGCAGGGCTCCGATGTATAATAATTGATGTCATAGATTGGACTTGACACTTGATAATCCATCTTGTTCCACCCTGTGCATAAATAAAAAGTGATCTTTTATAAAGTCCTAGAATGTATTTAGTTGCCCTCCATGAATGCAAACTGTTTTATACATCAATAGGTTTTTAATTGCCTACATAGATGTCTACATTGAATTAACTCTCTTTTTGGCCAAGCAATGAAGTTTTGTAGTGAAGGGAAGGTTTGCTGCTAGCTTCCCTGTCCACTAGATGGAGAGCTTGGCTCTGTTGGGGGAATTCATGAAAGCACCATCTCACCAAATAAAATCTTGTGCTCTATAGCACCATGGAGTGAATGAAGCTTTGACAACAATTAAGGGCGAATTCGCGGCCGCTAAATTCAATTCGCCCTATAGTGAGTCGTATTACAATTCACTGGCCGTCGTTTTACAACGTCGTGACTGGGAAAACCCTGGCGTTACCCAACTTAATCGCCTTGCAGCACATCCCCCTTTCGCCAGCTGGCGTAATAGCGAAGAGGCCCGCACCGATCGCCCTTCCCAACAGTTGCGCAGCCTATACGTACGGCAGTTTAAGGTTTACACCTATAAAAGAGAGAGCCGTTATCGTCTGTTTGTGGATGTACAGAGTGATATTATTGACACGCCGGGGCGACGGATGGTGATCCCCCTGGCCAGTGCACGTCTGCTGTCAGATAAAGTCTCCCGTGAACTTTACCCGGTGGTGCATATCGGGGATGAAAGCTGGCGCATGATGACCACCGATATGGCCAGTGTGCCGGTCTCCGTTATCGGGGAAGAAGTGGCTGATCTCAGCCACCGCGAAAATGACATCAAAAACGCCATTAACCTGATGTTCTGGGGAATATAAATGTCAGGCATGAGATTATCAAAAAGGATCTTCACCTAGATCCTTTTCACGTAGAAAGCCAGTCCGCAGAAACGGTGCTGACCCCGGATGAATGTCAGCTACTGGGCTATCTGGACAAGGGAAAACGCAAGCGCAAAGAGAAAGCAGGTAGCTTGCAGTGGGCTTACATGGCGATAGCTAGACTGGGCGGTTTTATGGACAGCAAGCGAACCGGAATTGCCAGCTGGGGCGCCCTCTGGTAAGGTTGGGAAGCCCTGCAAAGTAAACTGGATGGCTTTCTTGCCGCCAAGGATCTGATGGCGCAGGGGATCAAGCTCTGATCAAGAGACAGGATGAGGATCGTTTCGCATGATTGAACAAGATGGATTGCACGCAGGTTCTCCGGCCGCTTGGGTGGAGAGGCTATTCGGCTATGACTGGGCACAACAGACAATCGGCTGCTCTGATGCCGCCGTGTTCCGGCTGTCAGCGCAGGGGCGCCCGGTTCTTTTTGTCAAGACCGACCTGTCCGGTGCCCTGAATGAACTGCAAGACGAGGCAGCGCGGCTATCGTGGCTGGCCACGACGGGCGTTCCTTGCGCAGCTGTGCTCGACGTTGTCACTGAAGCGGGAAGGGACTGGCTGCTATTGGGCGAAGTGCCGGGGCAGGATCTCCTGTCATCTCACCTTGCTCCTGCCGAGAAAGTATCCATCATGGCTGATGCAATGCGGCGGCTGCATACGCTTGATCCGGCTACCTGCCCATTCGACCACCAAGCGAAACATCGCATCGAGCGAGCACGTACTCGGATGGAAGCCGGTCTTGTCGATCAGGATGATCTGGACGAAGAGCATCAGGGGCTCGCGCCAGCCGAACTGTTCGCCAGGCTCAAGGCGAGCATGCCCGACGGCGAGGATCTCGTCGTGACCCATGGCGATGCCTGCTTGCCGAATATCATGGTGGAAAATGGCCGCTTTTCTGGATTCATCGACTGTGGCCGGCTGGGTGTGGCGGACCGCTATCAGGACATAGCGTTGGCTACCCGTGATATTGCTGAAGAGCTTGGCGGCGAATGGGCTGACCGCTTCCTCGTGCTTTACGGTATCGCCGCTCCCGATTCGCAGCGCATCGCCTTCTATCGCCTTCTTGACGAGTTCTTCTGAATTATTAACGCTTACAATTTCCTGATGCGGTATTTTCTCCTTACGCATCTGTGCGGTATTTCACACCGCATCAGGTGGCACTTTTCGGGGAAATGTGCGCGGAACCCCTATTTGTTTATTTTTCTAAATACATTCAAATATGTATCCGCTCATGAGATTATCAAAAAGGATCTTCACCTAGATCCTTTTAAATTAAAAATGAAGTTTTAAATCAATCTAAAGTATATATGAGTAAACTTGGTCTGACAGTTACCAATGCTTAATCAGTGAGGCACCTATCTCAGCGATCTGTCTATTTCGTTCATCCATAGTTGCCTGACTCCCCGTCGTGTAGATAACTACGATACGGGAGGGCTTACCATCTGGCCCCAGTGCTGCAATGATACCGCGAGACCCACGCTCACCGGCTCCAGATTTATCAGCAATAAACCAGCCAGCCGGAAGGGCCGAGCGCAGAAGTGGTCCTGCAACTTTATCCGCCTCCATCCAGTCTATTAATTGTTGCCGGGAAGCTAGAGTAAGTAGTTCGCCAGTTAATAGTTTGCGCAACGTTGTTGCCATTGCTACAGGCATCGTGGTGTCACGCTCGTCGTTTGGTATGGCTTCATTCAGCTCCGGTTCCCAACGATCAAGGCGAGTTACATGATCCCCCATGTTGTGCAAAAAAGCGGTTAGCTCCTTCGGTCCTCCGATCGTTGTCAGAAGTAAGTTGGCCGCAGTGTTATCACTCATGGTTATGGCAGCACTGCATAATTCTCTTACTGTCATGCCATCCGTAAGATGCTTTTCTGTGACTGGTGAGTACTCAACCAAGTCATTCTGAGAATAGTGTATGCGGCGACCGAGTTGCTCTTGCCCGGCGTCAATACGGGATAATACCGCGCCACATAGCAGAACTTTAAAAGTGCTCATCATTGGAAAACGTTCTTCGGGGCGAAAACTCTCAAGGATCTTACCGCTGTTGAGATCCAGTTCGATGTAACCCACTCGTGCACCCAACTGATCTTCAGCATCTTTTACTTTCACCAGCGTTTCTGGGTGAGCAAAAACAGGAAGGCAAAATGCCGCAAAAAAGGGAATAAGGGCGACACGGAAATGTTGAATACTCATACTCTTCCTTTTTCAATATTATTGAAGCATTTATCAGGGTTATTGTCTCATGACCAAAATCCCTTAACGTGAGTTTTCGTTCCACTGAGCGTCAGACCCCGTAGAAAAGATCAAAGGATCTTCTTGAGATCCTTTTTTTCTGCGCGTAATCTGCTGCTTGCAAACAAAAAAACCACCGCTACCAGCGGTGGTTTGTTTGCCGGATCAAGAGCTACCAACTCTTTTTCCGAAGGTAACTGGCTTCAGCAGAGCGCAGATACCAAATACTGTTCTTCTAGTGTAGCCGTAGTTAGGCCACCACTTCAAGAACTCTGTAGCACCGCCTACATACCTCGCTCTGCTAATCCTGTTACCAGTGGCTGCTGCCAGTGGCGATAAGTCGTGTCTTACCGGGTTGGACTCAAGACGATAGTTACCGGATAAGGCGCAGCGGTCGGGCTGAACGGGGGGTTCGTGCACACAGCCCAGCTTGGAGCGAACGACCTACACCGAACTGAGATACCTACAGCGTGAGCTATGAGAAAGCGCCACGCTTCCCGAAGGGAGAAAGGCGGACAGGTATCCGGTAAGCGGCAGGGTCGGAACAGGAGAGCGCACGAGGGAGCTTCCAGGGGGAAACGCCTGGTATCTTTATAGTCCTGTCGGGTTTCGCCACCTCTGACTTGAGCGTCGATTTTTGTGATGCTCGTCAGGGGGGCGGAGCCTATGGAAAAACGCCAGCAACGCGGCCTTTTTACGGTTCCTGGCCTTTTGCTGGCCTTTTGCTCACATGTTCTTTCCTGCGTTATCCCCTGATTCTGTGGATAACCGTATTACCGCCTTTGAGTGAGCTGATACCGCTCGCCGCAGCCGAACGACCGAGCGCAGCGAGTCAGTGAGCGAGGAAGCGGAAG
GGTACCGAGCTCTTACGCGTGCTAGTATAAATACCTTCTGCCTTACTAGTATAAATACCTTCTGCCTTGCTAGCTCGAGATCTGCGATCTGCATCTCAATTAGTCAGCAACCATAGTCCCGCCCCTAACTCCGCCCATCCCGCCCCTAACTCCGCCCAGTTCCGCCCATTCTCCGCCCCATCGCTGACTAATTTTTTTTATTTATGCAGAGGCCGAGGCCGCCTCGGCCTCTGAGCTATTCCAGAAGTAGTGAGGAGGCTTTTTTGGAGGCCTAGGCTTTTGCAAAAAGCTTGGCATTCCGGTACTGTTGGTAAAGCCACCATGGAAGACGCCAAAAACATAAAGAAAGGCCCGGCGCCATTCTATCCGCTGGAAGATGGAACCGCTGGAGAGCAACTGCATAAGGCTATGAAGAGATACGCCCTGGTTCCTGGAACAATTGCTTTTACAGATGCACATATCGAGGTGGACATCACTTACGCTGAGTACTTCGAAATGTCCGTTCGGTTGGCAGAAGCTATGAAACGATATGGGCTGAATACAAATCACAGAATCGTCGTATGCAGTGAAAACTCTCTTCAATTCTTTATGCCGGTGTTGGGCGCGTTATTTATCGGAGTTGCAGTTGCGCCCGCGAACGACATTTATAATGAACGTGAATTGCTCAACAGTATGGGCATTTCGCAGCCTACCGTGGTGTTCGTTTCCAAAAAGGGGTTGCAAAAAATTTTGAACGTGCAAAAAAAGCTCCCAATCATCCAAAAAATTATTATCATGGATTCTAAAACGGATTACCAGGGATTTCAGTCGATGTACACGTTCGTCACATCTCATCTACCTCCCGGTTTTAATGAATACGATTTTGTGCCAGAGTCCTTCGATAGGGACAAGACAATTGCACTGATCATGAACTCCTCTGGATCTACTGGTCTGCCTAAAGGTGTCGCTCTGCCTCATAGAACTGCCTGCGTGAGATTCTCGCATGCCAGAGATCCTATTTTTGGCAATCAAATCATTCCGGATACTGCGATTTTAAGTGTTGTTCCATTCCATCACGGTTTTGGAATGTTTACTACACTCGGATATTTGATATGTGGATTTCGAGTCGTCTTAATGTATAGATTTGAAGAAGAGCTGTTTCTGAGGAGCCTTCAGGATTACAAGATTCAAAGTGCGCTGCTGGTGCCAACCCTATTCTCCTTCTTCGCCAAAAGCACTCTGATTGACAAATACGATTTATCTAATTTACACGAAATTGCTTCTGGTGGCGCTCCCCTCTCTAAGGAAGTCGGGGAAGCGGTTGCCAAGAGGTTCCATCTGCCAGGTATCAGGCAAGGATATGGGCTCACTGAGACTACATCAGCTATTCTGATTACACCCGAGGGGGATGATAAACCGGGCGCGGTCGGTAAAGTTGTTCCATTTTTTGAAGCGAAGGTTGTGGATCTGGATACCGGGAAAACGCTGGGCGTTAATCAAAGAGGCGAACTGTGTGTGAGAGGTCCTATGATTATGTCCGGTTATGTAAACAATCCGGAAGCGACCAACGCCTTGATTGACAAGGATGGATGGCTACATTCTGGAGACATAGCTTACTGGGACGAAGACGAACACTTCTTCATCGTTGACCGCCTGAAGTCTCTGATTAAGTACAAAGGCTATCAGGTGGCTCCCGCTGAATTGGAATCCATCTTGCTCCAACACCCCAACATCTTCGACGCAGGTGTCGCAGGTCTTCCCGACGATGACGCCGGTGAACTTCCCGCCGCCGTTGTTGTTTTGGAGCACGGAAAGACGATGACGGAAAAAGAGATCGTGGATTACGTCGCCAGTCAAGTAACAACCGCGAAAAAGTTGCGCGGAGGAGTTGTGTTTGTGGACGAAGTACCGAAAGGTCTTACCGGAAAACTCGACGCAAGAAAAATCAGAGAGATCCTCATAAAGGCCAAGAAGGGCGGAAAGATCGCCGTGTAATTCTAGAGTCGGGGCGGCCGGCCGCTTCGAGCAGACATGATAAGATACATTGATGAGTTTGGACAAACCACAACTAGAATGCAGTGAAAAAAATGCTTTATTTGTGAAATTTGTGATGCTATTGCTTTATTTGTAACCATTATAAGCTGCAATAAACAAGTTAACAACAACAATTGCATTCATTTTATGTTTCAGGTTCAGGGGGAGGTGTGGGAGGTTTTTTAAAGCAAGTAAAACCTCTACAAATGTGGTAAAATCGATAAGGATCCGTCGACCGATGCCCTTGAGAGCCTTCAACCCAGTCAGCTCCTTCCGGTGGGCGCGGGGCATGACTATCGTCGCCGCACTTATGACTGTCTTCTTTATCATGCAACTCGTAGGACAGGTGCCGGCAGCGCTCTTCCGCTTCCTCGCTCACTGACTCGCTGCGCTCGGTCGTTCGGCTGCGGCGAGCGGTATCAGCTCACTCAAAGGCGGTAATACGGTTATCCACAGAATCAGGGGATAACGCAGGAAAGAACATGTGAGCAAAAGGCCAGCAAAAGGCCAGGAACCGTAAAAAGGCCGCGTTGCTGGCGTTTTTCCATAGGCTCCGCCCCCCTGACGAGCATCACAAAAATCGACGCTCAAGTCAGAGGTGGCGAAACCCGACAGGACTATAAAGATACCAGGCGTTTCCCCCTGGAAGCTCCCTCGTGCGCTCTCCTGTTCCGACCCTGCCGCTTACCGGATACCTGTCCGCCTTTCTCCCTTCGGGAAGCGTGGCGCTTTCTCATAGCTCACGCTGTAGGTATCTCAGTTCGGTGTAGGTCGTTCGCTCCAAGCTGGGCTGTGTGCACGAACCCCCCGTTCAGCCCGACCGCTGCGCCTTATCCGGTAACTATCGTCTTGAGTCCAACCCGGTAAGACACGACTTATCGCCACTGGCAGCAGCCACTGGTAACAGGATTAGCAGAGCGAGGTATGTAGGCGGTGCTACAGAGTTCTTGAAGTGGTGGCCTAACTACGGCTACACTAGAAGAACAGTATTTGGTATCTGCGCTCTGCTGAAGCCAGTTACCTTCGGAAAAAGAGTTGGTAGCTCTTGATCCGGCAAACAAACCACCGCTGGTAGCGGTGGTTTTTTTGTTTGCAAGCAGCAGATTACGCGCAGAAAAAAAGGATCTCAAGAAGATCCTTTGATCTTTTCTACGGGGTCTGACGCTCAGTGGAACGAAAACTCACGTTAAGGGATTTTGGTCATGAGATTATCAAAAAGGATCTTCACCTAGATCCTTTTAAATTAAAAATGAAGTTTTAAATCAATCTAAAGTATATATGAGTAAACTTGGTCTGACAGTTACCAATGCTTAATCAGTGAGGCACCTATCTCAGCGATCTGTCTATTTCGTTCATCCATAGTTGCCTGACTCCCCGTCGTGTAGATAACTACGATACGGGAGGGCTTACCATCTGGCCCCAGTGCTGCAATGATACCGCGAGACCCACGCTCACCGGCTCCAGATTTATCAGCAATAAACCAGCCAGCCGGAAGGGCCGAGCGCAGAAGTGGTCCTGCAACTTTATCCGCCTCCATCCAGTCTATTAATTGTTGCCGGGAAGCTAGAGTAAGTAGTTCGCCAGTTAATAGTTTGCGCAACGTTGTTGCCATTGCTACAGGCATCGTGGTGTCACGCTCGTCGTTTGGTATGGCTTCATTCAGCTCCGGTTCCCAACGATCAAGGCGAGTTACATGATCCCCCATGTTGTGCAAAAAAGCGGTTAGCTCCTTCGGTCCTCCGATCGTTGTCAGAAGTAAGTTGGCCGCAGTGTTATCACTCATGGTTATGGCAGCACTGCATAATTCTCTTACTGTCATGCCATCCGTAAGATGCTTTTCTGTGACTGGTGAGTACTCAACCAAGTCATTCTGAGAATAGTGTATGCGGCGACCGAGTTGCTCTTGCCCGGCGTCAATACGGGATAATACCGCGCCACATAGCAGAACTTTAAAAGTGCTCATCATTGGAAAACGTTCTTCGGGGCGAAAACTCTCAAGGATCTTACCGCTGTTGAGATCCAGTTCGATGTAACCCACTCGTGCACCCAACTGATCTTCAGCATCTTTTACTTTCACCAGCGTTTCTGGGTGAGCAAAAACAGGAAGGCAAAATGCCGCAAAAAAGGGAATAAGGGCGACACGGAAATGTTGAATACTCATACTCTTCCTTTTTCAATATTATTGAAGCATTTATCAGGGTTATTGTCTCATGAGCGGATACATATTTGAATGTATTTAGAAAAATAAACAAATAGGGGTTCCGCGCACATTTCCCCGAAAAGTGCCACCTGACGCGCCCTGTAGCGGCGCATTAAGCGCGGCGGGTGTGGTGGTTACGCGCAGCGTGACCGCTACACTTGCCAGCGCCCTAGCGCCCGCTCCTTTCGCTTTCTTCCCTTCCTTTCTCGCCACGTTCGCCGGCTTTCCCCGTCAAGCTCTAAATCGGGGGCTCCCTTTAGGGTTCCGATTTAGTGCTTTACGGCACCTCGACCCCAAAAAACTTGATTAGGGTGATGGTTCACGTAGTGGGCCATCGCCCTGATAGACGGTTTTTCGCCCTTTGACGTTGGAGTCCACGTTCTTTAATAGTGGACTCTTGTTCCAAACTGGAACAACACTCAACCCTATCTCGGTCTATTCTTTTGATTTATAAGGGATTTTGCCGATTTCGGCCTATTGGTTAAAAAATGAGCTGATTTAACAAAAATTTAACGCGAATTTTAACAAAATATTAACGCTTACAATTTGCCATTCGCCATTCAGGCTGCGCAACTGTTGGGAAGGGCGATCGGTGCGGGCCTCTTCGCTATTACGCCAGCCCAAGCTACCATGATAAGTAAGTAATATTAAGGTACGGGAGGTACTTGGAGCGGCCGCAATAAAATATCTTTATTTTCATTACATCTGTGTGTTGGTTTTTTGTGTGAATCGATAGTACTAACATACGCTCTCCATCAAAACAAAACGAAACAAAACAAACTAGCAAAATAGGCTGTCCCCAGTGCAAGTGCAGGTGCCAGAACATTTCTCTATCGATA
GTGGATCTACGCACGCTCGGCTACAGCCAGCAGCAACAGGAGAAGATCAAACCGAAGGTTCGTTCGACAGTGGCGCAGCACCACGAGGCACTGGTCGGCCATGGGTTTACACACGCGCACATCGTTGCGCTCAGCCAACACCCGGCAGCGTTAGGGACCGTCGCTGTCAAGTATCAGGACATGATCGCAGCGTTGCCAGAGGCGACACACGAAGCGATCGTTGGCGTCGGCAAACAGTGGTCCGGCGCACGCGCCCTGGAGGCCTTGCTCACGGTGGCGGGAGAGTTGAGAGGTCCACCGTTACAGTTGGACACAGGCCAACTTCTCAAGATTGCAAAACGTGGCGGCGTGACCGCAGTGGAGGCAGTGCATGCATGGCGCAATGCACTGACGGGTGCCCCCCTGAACCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCAATATTGGTGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCAATGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCAATATTGGTGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCAATATTGGTGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGCACAGGTGGTGGCCATCGCCAGCAATATTGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGCGCCAGGCCCATGGCCTGACCCCGGAGCAGGTCGTGGCCATCGCCAGCAATGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGCACAGGTGGTGGCCATCGCCAGCAATATTGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTCGTGGCCATCGCCAGCCACGATGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCCACGATGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGACCAGGTGGTGGCCATCGCCAGCAATGGCGGTGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGACCAGGTGGTGGCCATCGCCAGCAATGGCGGTGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCCACGATGGCGGCAAGCAGGCGCTGGAGACGGTGCAGCGGCTGTTGCCGGTGCTGTGCCAGGCCCATGGCCTGACCCCGGAGCAGGTGGTGGCCATCGCCAGCAATGGCGGCGGCAGGCCGGCGCTGGAGAGCATTGTTGCCCAGTTATCTCGCCCTGATCCGGCGTTGGCCGCGTTGACCAACGACCACCTCGTCGCCTTGGCCTGCCTCGGCGGACGTCCTGCGCTGGATGCAGTGAAAAAGGGATTGCCGCACGCGCCGGCCTTGATCAAAAGAACCAATCGCCGTATTCCCGAACGCACATCCCATCGCGTTGCCGACCACGCGCAAGTGGTTCGCGTGCTGGGTTTTTTCCAGTGCCACTCCCACCCAGCGCAAGCATTTGATGACGCCATGACGCAGTTCGGGATGAGCAGGCACGGGTTGTTACAGCTCTTTCGCAGAGTGGGCGTCACCGAACTCGAAGCCCGCAGTGGAACGCTCCCCCCAGCCTCGCAGCGTTGGGACCGTATCCTCCAGGCATCAGGGATGAAAAGGGCCAAACCGTCCCCTACTTCAACTCAAACGCCGGACCAGGCGTCTTTGCATGCATTCGCCGATTCGCTGGAGCGTGACCTTGATGCGCCCAGCCCAACGCACGAGGGAGATCAGAGGCGGGCAAGCAGCCGTAAACGGTCCCGATCGGATCGTGCTGTCACCGGTCCCTCCGCACAGCAATCGTTCGAGGTGCGCGCTCCCGAACAGCGCGATGCGCTGCATTTGCCCCTCAGTTGGAGGGTAAAACGCCCGCGTACCAGTATCGGGGGCGGCCTCCCGGATCCTGGTACGCCCACGGCTGCCGACCTGGCAGCGTCCAGCACCGTGATGCGGGAACAAGATGAGGACCCCTTCGCAGGGGCAGCGGATGATTTCCCGGCATTCAACGAAGAGGAGCTCGCATGGTTGATGGAGCTATTGCCTCAG
>VEGF−1(配列番号321)
VDLRTLGYSQQQQEKIKPKVRSTVAQHHEALVGHGFTHAHIVALSQHPAALGTVAVKYQDMIAALPEATHEAIVGVGKQWSGARALEALLTVAGELRGPPLQLDTGQLLKIAKRGGVTAVEAVHAWRNALTGAPLNLTPQQVVAIASNNGGKQALETVQRLLPVLCQAHGLTPQQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNNGGKQALETVQRLLPVLCQAHGLTPEQVVAIASHDGGKQALETVQALLPVLCQAHGLTPEQVVAIASHDGGKQALETVQALLPVLCQAHGLTPEQVVAIASHDGGKQALETVQALLPVLCQAHGLTPEQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPQQVVAIASNGGGRPALESIVAQLSRPDPALAALTNDHLVALACLGGRPALDAVKKGLPHAPALIKRTNRRIPERTSHRVADHAQVVRVLGFFQCHSHPAQAFDDAMTQFGMS
GTGGATCTACGCACGCTCGGCTACAGCCAGCAGCAACAGGAGAAGATCAAACCGAAGGTTCGTTCGACAGTGGCGCAGCACCACGAGGCACTGGTCGGCCATGGGTTTACACACGCGCACATCGTTGCGCTCAGCCAACACCCGGCAGCGTTAGGGACCGTCGCTGTCAAGTATCAGGACATGATCGCAGCGTTGCCAGAGGCGACACACGAAGCGATCGTTGGCGTCGGCAAACAGTGGTCCGGCGCACGCGCCCTGGAGGCCTTGCTCACGGTGGCGGGAGAGTTGAGAGGTCCACCGTTACAGTTGGACACAGGCCAACTTCTCAAGATTGCAAAACGTGGCGGCGTGACCGCAGTGGAGGCAGTGCATGCATGGCGCAATGCACTGACGGGGGCCCCCCTGAACCTGACGCCTCAACAGGTCGTCGCGATAGCGTCTAATAATGGAGGAAAGCAAGCTCTGGAAACCGTCCAGCGACTCCTTCCGGTTCTGTGCCAGGCTCATGGTCTGACTCCGCAGCAAGTCGTTGCTATAGCGTCCAACATCGGAGGCAAACAGGCCCTGGAGACCGTGCAGCGGTTGTTGCCTGTGCTTTGCCAAGCCCACGGGCTTACGCCTGAGCAAGTGGTGGCGATTGCCAGTAACAACGGCGGCAAACAAGCCCTTGAGACTGTGCAGAGGCTCTTGCCGGTACTCTGCCAAGCACACGGCTTGACCCCCGAGCAGGTTGTAGCCATAGCTAGTCACGACGGGGGTAAGCAAGCGTTGGAAACGGTGCAAGCACTTCTCCCCGTTCTCTGTCAAGCGCATGGACTTACCCCGGAACAGGTGGTCGCCATTGCAAGCCATGATGGGGGTAAGCAAGCGTTGGAAACGGTGCAAGCACTTCTCCCCGTTCTCTGTCAAGCGCATGGACTTACCCCGGAACAGGTGGTCGCCATTGCAAGCCATGATGGAGGAAAGCAGGCGCTCGAAACAGTCCAGGCACTTTTGCCCGTACTTTGTCAAGCTCACGGTCTCACCCCGGAACAGGTGGTAGCCATTGCATCTAACGGAGGGGGCAAACAAGCCTTGGAGACAGTGCAAAGGCTCCTGCCAGTGCTCTGCCAGGCTCATGGTTTGACACCCGAACAGGTAGTTGCAATAGCGAGTCATGATGGCGGAAAGCAAGCTCTTGAAACTGTGCAGCGGCTGTTGCCTGTACTGTGTCAAGCCCACGGGCTGACACCGGAACAAGTTGTAGCGATCGCTAGCCACGATGGCGGGAAACAAGCTCTGGAAACGGTACAGAGACTCCTCCCAGTGCTTTGTCAGGCACACGGCCTCACGCCAGAGCAGGTTGTCGCCATCGCGTCACATGATGGGGGCAAACAAGCCTTGGAGACAGTGCAAAGGCTCCTGCCAGTGCTCTGCCAGGCTCATGGTTTGACACCCGAACAGGTAGTTGCAATAGCGAGTCATGATGGCGGAAAGCAAGCTCTTGAAACTGTGCAGCGGCTGTTGCCTGTACTGTGTCAAGCCCACGGGCTGACACCGGAACAAGTTGTAGCGATCGCTAGCCACGATGGCGGGAAACAAGCTCTGGAAACGGTACAGAGACTCCTCCCAGTGCTTTGTCAGGCACACGGCCTCACGCCAGAGCAGGTTGTCGCCATCGCGTCAAACGGTGGAGGGAAACAAGCGCTCGAAACCGTGCAAAGGTTGCTCCCCGTTCTCTGTCAGGCGCACGGTCTTACGCCACAACAGGTGGTGGCGATTGCATCTAATGGAGGCGGACGCCCTGCCTTGGAGAGCATTGTGGCCCAGCTGTCCAGGCCGGACCCTGCCCTGGCCGCGTTAACCAACGACCACCTCGTCGCCTTGGCCTGCCTCGGCGGACGTCCTGCGCTGGATGCAGTGAAAAAGGGATTGCCGCACGCGCCGGCCTTGATCAAAAGAACCAATCGCCGTATTCCCGAACGCACATCCCATCGCGTTGCCGACCACGCGCAAGTGGTTCGCGTGCTGGGTTTTTTCCAGTGCCACTCCCACCCAGCGCAAGCATTTGATGACGCCATGACGCAGTTCGGGATGAGC
VDLRTLGYSQQQQEKIKPKVRSTVAQHHEALVGHGFTHAHIVALSQHPAALGTVAVKYQDMIAALPEATHEAIVGVGKQWSGARALEALLTVAGELRGPPLQLDTGQLLKIAKRGGVTAVEAVHAWRNALTGAPLNLTPEQVVAIASNKGGKQALETVQALLPVLCQAHGLTPEQVVAIASHDGGKQALETVQALLPVLCQAHGLTPEQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNGGGKQALETVQALLPVLCQAHGLTPEQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPQQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNNGGKQALETVQRLLPVLCQAHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPEQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPQQVVAIASNGGGRPALESIVAQLSRPDPALAALTNDHLVALACLGGRPALDAVKKGLPHAPALIKRTNRRIPERTSHRVADHAQVVRVLGFFQCHSHPAQAFDDAMTQFGMS
GTGGATCTACGCACGCTCGGCTACAGCCAGCAGCAACAGGAGAAGATCAAACCGAAGGTTCGTTCGACAGTGGCGCAGCACCACGAGGCACTGGTCGGCCATGGGTTTACACACGCGCACATCGTTGCGCTCAGCCAACACCCGGCAGCGTTAGGGACCGTCGCTGTCAAGTATCAGGACATGATCGCAGCGTTGCCAGAGGCGACACACGAAGCGATCGTTGGCGTCGGCAAACAGTGGTCCGGCGCACGCGCCCTGGAGGCCTTGCTCACGGTGGCGGGAGAGTTGAGAGGTCCACCGTTACAGTTGGACACAGGCCAACTTCTCAAGATTGCAAAACGTGGCGGCGTGACCGCAGTGGAGGCAGTGCATGCATGGCGCAATGCACTGACGGGGGCCCCCCTGAACCTTACACCCGAGCAAGTAGTGGCTATTGCGAGTAATAAAGGGGGTAAGCAAGCGTTGGAAACGGTGCAAGCACTTCTCCCCGTTCTCTGTCAAGCGCATGGACTTACCCCGGAACAGGTGGTCGCCATTGCAAGCCATGATGGAGGAAAGCAGGCGCTCGAAACAGTCCAGGCACTTTTGCCCGTACTTTGTCAAGCTCACGGTCTCACCCCGGAACAGGTGGTAGCCATTGCATCTAACGGAGGGGGCAAACAAGCCTTGGAGACAGTGCAAAGGCTCCTGCCAGTGCTCTGCCAGGCTCATGGTTTGACACCCGAACAGGTAGTTGCAATAGCGAGTCATGATGGCGGAAAGCAAGCTCTTGAAACTGTGCAGCGGCTGTTGCCTGTACTGTGTCAAGCCCACGGGCTGACACCGGAACAAGTTGTAGCGATCGCTAGCAACGGCGGAGGTAAGCAAGCATTGGAAACGGTTCAGGCCCTGTTGCCTGTACTTTGCCAGGCGCACGGTCTGACACCTGAGCAGGTTGTCGCCATCGCTAGCAACGGAGGTGGGAAACAGGCACTTGAAACTGTGCAGAGGCTTCTGCCGGTGCTGTGCCAAGCGCATGGCCTTACACCCGAGCAAGTAGTGGCTATTGCGAGTCATGATGGAGGCAAGCAAGCGCTGGAGACTGTCCAACGACTTCTTCCGGTCTTGTGTCAGGCACATGGATTGACCCCTCAACAAGTCGTGGCGATAGCTAGCAACATCGGAGGCAAACAGGCCCTGGAGACCGTGCAGCGGTTGTTGCCTGTGCTTTGCCAAGCCCACGGGCTTACGCCTGAGCAAGTGGTGGCGATTGCCAGTAACAACGGGGGCAAACAAGCCTTGGAGACAGTGCAAAGGCTCCTGCCAGTGCTCTGCCAGGCTCATGGTTTGACACCCGAACAGGTAGTTGCAATAGCGAGTCATGATGGCGGAAAGCAAGCTCTTGAAACTGTGCAGCGGCTGTTGCCTGTACTGTGTCAAGCCCACGGGCTGACACCGGAACAAGTTGTAGCGATCGCTAGCCACGATGGCGGGAAACAAGCTCTGGAAACGGTACAGAGACTCCTCCCAGTGCTTTGTCAGGCACACGGCCTCACGCCAGAGCAGGTTGTCGCCATCGCGTCAAACGGTGGAGGGAAACAAGCGCTCGAAACCGTGCAAAGGTTGCTCCCCGTTCTCTGTCAGGCGCACGGTCTTACGCCACAACAGGTGGTGGCGATTGCATCTAATGGAGGCGGACGCCCTGCCTTGGAGAGCATTGTGGCCCAGCTGTCCAGGCCGGACCCTGCCCTGGCCGCGTTAACCAACGACCACCTCGTCGCCTTGGCCTGCCTCGGCGGACGTCCTGCGCTGGATGCAGTGAAAAAGGGATTGCCGCACGCGCCGGCCTTGATCAAAAGAACCAATCGCCGTATTCCCGAACGCACATCCCATCGCGTTGCCGACCACGCGCAAGTGGTTCGCGTGCTGGGTTTTTTCCAGTGCCACTCCCACCCAGCGCAAGCATTTGATGACGCCATGACGCAGTTCGGGATGAGC
101077ORF(下線部分はTALE領域)(配列番号325):
MDYKDHDGDYKDHDIDYKDDDDKMAPKKKRKVGIHRGVPMVDLRTLGYSQQQQEKIKPKVRSTVAQHHEALVGHGFTHAHIVALSQHPAALGTVAVKYQDMIAALPEATHEAIVGVGKQWSGARALEALLTVAGELRGPPLQLDTGQLLKIAKRGGVTAVEAVHAWRNALTGAPLNLTPDQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPAQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQDHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPAQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPDQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPAQVVAIANNNGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGRPALESIVAQLSRPDPALAALTNDHLVALACLGGRPALDAVKKGLPHAPALIKRTNRRIPERTSHRVAGSQLVKSELEEKKSELRHKLKYVPHEYIELIEIARNSTQDRILEMKVMEFFMKVYGYRGKHLGGSRKPDGAIYTVGSPIDYGVIVDTKAYSGGYNLPIGQADEMQRYVEENQTRNKHINPNEWWKVYPSSVTEFKFLFVSGHFKGNYKAQLTRLNHITNCNGAVLSVEELLIGGEMIKAGTLTLEEVRRKFNNGEINFRS
MDYKDHDGDYKDHDIDYKDDDDKMAPKKKRKVGIHRGVPMVDLRTLGYSQQQQEKIKPKVRSTVAQHHEALVGHGFTHAHIVALSQHPAALGTVAVKYQDMIAALPEATHEAIVGVGKQWSGARALEALLTVAGELRGPPLQLDTGQLLKIAKRGGVTAVEAVHAWRNALTGAPLNLTPDQVVAIASNGGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPAQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQDHGLTPDQVVAIANNNGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPAQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPDQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPDQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPAQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGRPALESIVAQLSRPDPALAALTNDHLVALACLGGRPALDAVKKGLPHAPALIKRTNRRIPERTSHRVAGSQLVKSELEEKKSELRHKLKYVPHEYIELIEIARNSTQDRILEMKVMEFFMKVYGYRGKHLGGSRKPDGAIYTVGSPIDYGVIVDTKAYSGGYNLPIGQADEMQRYVEENQTRNKHINPNEWWKVYPSSVTEFKFLFVSGHFKGNYKAQLTRLNHITNCNGAVLSVEELLIGGEMIKAGTLTLEEVRRKFNNGEINFRS
101318(配列番号327)
MDYKDHDGDYKDHDIDYKDDDDKMAPKKKRKVGIHRGVPMVDLRTLGYSQQQQEKIKPKVRSTVAQHHEALVGHGFTHAHIVALSQHPAALGTVAVKYQDMIAALPEATHEAIVGVGKQWSGARALEALLTVAGELRGPPLQLDTGQLLKIAKRGGVTAVEAVHAWRNALTGAPLNLTPDQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPAQVVAIANNNGGKQALETVQRLLPVLCQDHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPAQVVAIANNNGGKQALETVQRLLPVLCQDHGLTPDQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPDQVVAIANNNGGKQALETVQRLLPVLCQAHGLTPAQVVAIANNNGGKQALETVQRLLPVLCQDHGLTPDQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPAQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGRPALESIVAQLSRPDPALAALTNDHLVALACLGGRPALDAVKKGLPHAPALIKRTNRRIPERTSHRVAGSQLVKSELEEKKSELRHKLKYVPHEYIELIEIARNSTQDRILEMKVMEFFMKVYGYRGKHLGGSRKPDGAIYTVGSPIDYGVIVDTKAYSGGYNLPIGQADEMQRYVEENQTRNKHINPNEWWKVYPSSVTEFKFLFVSGHFKGNYKAQLTRLNHITNCNGAVLSVEELLIGGEMIKAGTLTLEEVRRKFNNGEINFRS
MDYKDHDGDYKDHDIDYKDDDDKMAPKKKRKVGIHRGVPMVDLRTLGYSQQQQEKIKPKVRSTVAQHHEALVGHGFTHAHIVALSQHPAALGTVAVKYQDMIAALPEATHEAIVGVGKQWSGARALEALLTVAGELRGPPLQLDTGQLLKIAKRGGVTAVEAVHAWRNALTGAPLNLTPDQVVAIASNGGGKQALETVQRLLPVLCQDHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNGGGKQALETVQRLLPVLCQDHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQDHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPDQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPEQVVAIASHDGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNGGGKQALETVQRLLPVLCQAHGLTPAQVVAIANNNGGKQALETVQRLLPVLCQDHGLTPDQVVAIASHDGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPDQVVAIASNIGGKQALETVQRLLPVLCQAHGLTPAQVVAIASNIGGKQALETVQRLLPVLCQDHGLTPEQVVAIASNGGGRPALESIVAQLSRPDPALAALTNDHLVALACLGGRPALDAVKKGLPHAPALIKRTNRRIPERTSHRVAGSQLVKSELEEKKSELRHKLKYVPHEYIELIEIARNSTQDRILEMKVMEFFMKVYGYRGKHLGGSRKPDGAIYTVGSPIDYGVIVDTKAYSGGYNLPIGQADEMQRYVEENQTRNKHINPNEWWKVYPSSVTEFKFLFVSGHFKGNYKAQLTRLNHITNCNGAVLSVEELLIGGEMIKAGTLTLEEVRRKFNNGEINFRS
ctttcctgcgttatcccctgattctgtggataaccgtattaccgcctttgagtgagctgataccgctcgccgcagccgaacgaccgagcgcagcgagtcagtgagcgaggaagcggaagagcgcccaatacgcaaaccgcctctccccgcgcgttggccgattcattaatgcagctggcacgacaggtttcccgactggaaagcgggcagtgagcgcaacgcaattaatacgcgtaccgctagccaggaagagtttgtagaaacgcaaaaaggccatccgtcaggatggccttctgcttagtttgatgcctggcagtttatggcgggcgtcctgcccgccaccctccgggccgttgcttcacaacgttcaaatccgctcccggcggatttgtcctactcaggagagcgttcaccgacaaacaacagataaaacgaaaggcccagtcttccgactgagcctttcgttttatttgatgcctggcagttccctactctcgcgttaacgctagcatggatgttttcccagtcacgacgttgtaaaacgacggccagtcttaagctcgggccccaaataatgattttattttgactgatagtgacctgttcgttgcaacaaattgatgagcaatgcttttttataatgccaactttgtacaaaaaagcaggctccgaattcgcccttttaattaatgcagtgcagcgtgacccggtcgtgcccctctctagagataatgagcattgcatgtctaagttataaaaaattaccacatattttttttgtcacacttgtttgaagtgcagtttatctatctttatacatatatttaaactttactctacgaataatataatctatagtactacaataatatcagtgttttagagaatcatataaatgaacagttagacatggtctaaaggacaattgagtattttgacaacaggactctacagttttatctttttagtgtgcatgtgttctcctttttttttgcaaatagcttcacctatataatacttcatccattttattagtacatccatttagggtttagggttaatggtttttatagactaatttttttagtacatctattttattctattttagcctctaaattaagaaaactaaaactctattttagtttttttatttaataatttagatataaaatagaataaaataaagtgactaaaaattaaacaaataccctttaagaaattaaaaaaactaaggaaacatttttcttgtttcgagtagataatgccagcctgttaaacgccgtcgacgagtctaacggacaccaaccagcgaaccagcagcgtcgcgtcgggccaagcgaagcagacggcacggcatctctgtcgctgcctctggacccctctcgagagttccgctccaccgttggacttgctccgctgtcggcatccagaaattgcgtggcggagcggcagacgtgagccggcacggcaggcggcctcctcctcctctcacggcaccggcagctacgggggattcctttcccaccgctccttcgctttcccttcctcgcccgccgtaataaatagacaccccctccacaccctctttccccaacctcgtgttgttcggagcgcacacacacacaaccagatctcccccaaatccacccgtcggcacctccgcttcaaggtacgccgctcgtcctccccccccccccctctctaccttctctagatcggcgttccggtccatggttagggcccggtagttctacttctgttcatgtttgtgttagatccgtgtttgtgttagatccgtgctgctagcgttcgtacacggatgcgacctgtacgtcagacacgttctgattgctaacttgccagtgtttctctttggggaatcctgggatggctctagccgttccgcagacgggatcgatttcatgattttttttgtttcgttgcatagggtttggtttgcccttttcctttatttcaatatatgccgtgcacttgtttgtcgggtcatcttttcatgcttttttttgtcttggttgtgatgatgtggtctggttgggcggtcgttctagatcggagtagaattctgtttcaaactacctggtggatttattaattttggatctgtatgtgtgtgccatacatattcatagttacgaattgaagatgatggatggaaatatcgatctaggataggtatacatgttgatgcgggttttactgatgcatatacagagatgctttttgttcgcttggttgtgatgatgtggtgtggttgggcggtcgttcattcgttctagatcggagtagaatactgtttcaaactacctggtgtatttattaattttggaactgtatgtgtgtgtcatacatcttcatagttacgagtttaagatggatggaaatatcgatctaggataggtatacatgttgatgtgggttttactgatgcatatacatgatggcatatgcagcatctattcatatgctctaaccttgagtacctatctattataataaacaagtatgttttataattattttgatcttgatatacttggatgatggcatatgcagcagctatatgtggatttttttagccctgccttcatacgctatttatttgcttggtactgtttcttttgtcgatgctcaccctgttgtttggtgttacttctgcaggactagtccagtgtggtggaattcgccatggactacaaagaccatgacggtgattataaagatcatgacatcgattacaaggatgacgatgacaagatggcccccaagaagaagaggaaggtgggcattcacggggtacctatggtggacttgaggacactcggttattcgcaacagcaacaggagaaaatcaagcctaaggtcaggagcaccgtcgcgcaacaccacgaggcgcttgtggggcatggcttcactcatgcgcatattgtcgcgctttcacagcaccctgcggcgcttgggacggtggctgtcaaataccaagatatgattgcggccctgcccgaagccacgcacgaggcaattgtaggggtcggtaaacagtggtcgggagcgcgagcacttgaggcgctgctgactgtggcgggtgagcttagggggcctccgctccagctcgacaccgggcagctgctgaagatcgcgaagagagggggagtaacagcggtagaggcagtgcacgcctggcgcaatgcgctcaccggggcccccttgaacctgaccccagaccaggtagtcgcaatcgcgtcgcatgacgggggaaagcaagccctggaaaccgtgcaaaggttgttgccggtcctttgtcaagaccacggccttacaccggagcaagtcgtggccattgcatcacatgacggtggcaaacaggctcttgagacggttcagagacttctcccagttctctgtcaagcccacgggctgactcccgatcaagttgtagcgattgcgagcaatgggggagggaaacaagcattggagactgtccaacggctccttcccgtgttgtgtcaagcccacggtttgacgcctgcacaagtggtcgccatcgcctccaatattggcggtaagcaggcgctggaaacagtacagcgcctgctgcctgtactgtgccaggatcatggactcaccccagaccaggtagtcgcaatcgcgtcgcatgacgggggaaagcaagccctggaaaccgtgcaaaggttgttgccggtcctttgtcaagaccacggccttacaccggatcaagtcgtggccattgcaaataataacggtggcaaacaggctcttgagacggttcagagacttctcccagttctctgtcaagcccacgggctgactcccgatcaagttgtagcgattgcgagcaacatcggagggaaacaagcattggagactgtccaacggctccttcccgtgttgtgtcaagcccacggtttgacgcctgcacaagtggtcgccatcgcctcccacgacggcggtaagcaggcgctggaaacagtacagcgcctgctgcctgtactgtgccaggatcatgggctgaccccagaccaggtagtcgcaatcgccaacaataacgggggaaagcaagccctggaaaccgtgcaaaggttgttgccggtcctttgtcaagaccacggccttacaccggagcaagtcgtggccattgcatcaaatatcggtggcaaacaggctcttgagacggttcagagacttctcccagttctctgtcaagcccacgggctgactcccgatcaagttgtagcgattgcgaataacaatggagggaaacaagcattggagactgtccaacggctccttcccgtgttgtgtcaagcccacggtttgacgcctgcacaagtggtcgccatcgccaacaacaacggcggtaagcaggcgctggaaacagtacagcgcctgctgcctgtactgtgccaggatcatggtttgaccccagaccaggtagtcgcaatcgcgtcgaacattgggggaaagcaagccctggaaaccgtgcaaaggttgttgccggtcctttgtcaagaccacggccttacaccggatcaagtcgtggccattgcaaataataacggtggcaaacaggctcttgagacggttcagagacttctcccagttctctgtcaagcccacgggctgactcccgatcaagttgtagcgattgcgaataacaatggagggaaacaagcattggagactgtccaacggctccttcccgtgttgtgtcaagcccacggtttgacgcctgcacaagtggtcgccatcgcctccaatattggcggtaagcaggcgctggaaacagtacagcgcctgctgcctgtactgtgccaggatcatggcctgacacccgaacaggtggtcgccattgctagcaacgggggaggacggccagccttggagtccatcgtagcccaattgtccaggcccgatcccgcgttggctgcgttaacgaatgaccatctggtggcgttggcatgtcttggtggacgacccgcgctcgatgcagtcaaaaagggtctgcctcatgctcccgcattgatcaaaagaaccaaccggcggattcccgagagaacttcccatcgagtcgcgggatcccagctggttaaatcagaactcgaagaaaaaaagagcgagctgcggcataaactcaaatatgtccctcatgagtacatagaactgattgaaatcgcccgcaattccacccaggatcggattcttgaaatgaaagtgatggaattttttatgaaagtttacggctatcgcgggaagcaccttggggggtcgcggaagccggacggtgctatttacactgtcggttccccgatcgattatggcgtaattgttgacacgaaagcatattcgggtgggtataatcttcctattggtcaggctgatgagatgcagcggtacgttgaagagaatcagacgcggaacaagcatattaacccaaatgagtggtggaaggtgtatccatcatcggtcaccgaatttaagttcttgtttgtgtcgggccactttaaggggaactacaaggcccaacttaccaggttgaatcacataaccaactgtaacggagctgttctgtcagtagaagagctgttgataggcggggaaatgattaaagcaggtacattaacgttggaggaagtacgccgcaagtttaataacggcgagattaactttagatctgagacctgataaacaaacacacggtctcctcgagctcgcagatcgttcaacatctggcaataaagtttcttaagattgaatcctgttgccggtcttgcgatgattatcatataatttctgttgaattacgttaagcatgtaataattaacatgtaatgcatgacgttatttatgagatgggtttttatgattagagtcccgcaattatacatttaatacgcgatagaaaacaaaatatagcgcgcaaactaggataaattatcgcgcgcggtgtcatctatgttactagatccgataagcttaagggcgaattcgacccagctttcttgtacaaagttggcattataaaaaataattgctcatcaatttgttgcaacgaacaggtcactatcagtcaaaataaaatcattatttgccatccagctgatatcccctatagtgagtcgtattacatggtcatagctgtttcctggcagctctggcccgtgtctcaaaatctctgatgttacattgcacaagataaaaatatatcatcatgcctcctctagaccagccaggacagaaatgcctcgacttcgctgctgcccaaggttgccgggtgacgcacaccgtggaaacggatgaaggcacgaacccagtggacataagcctgttcggttcgtaagctgtaatgcaagtagcgtatgcgctcacgcaactggtccagaaccttgaccgaacgcagcggtggtaacggcgcagtggcggttttcatggcttgttatgactgtttttttggggtacagtctatgcctcgggcatccaagcagcaagcgcgttacgccgtgggtcgatgtttgatgttatggagcagcaacgatgttacgcagcagggcagtcgccctaaaacaaagttaaacatcatgagggaagcggtgatcgccgaagtatcgactcaactatcagaggtagttggcgtcatcgagcgccatctcgaaccgacgttgctggccgtacatttgtacggctccgcagtggatggcggcctgaagccacacagtgatattgatttgctggttacggtgaccgtaaggcttgatgaaacaacgcggcgagctttgatcaacgaccttttggaaacttcggcttcccctggagagagcgagattctccgcgctgtagaagtcaccattgttgtgcacgacgacatcattccgtggcgttatccagctaagcgcgaactgcaatttggagaatggcagcgcaatgacattcttgcaggtatcttcgagccagccacgatcgacattgatctggctatcttgctgacaaaagcaagagaacatagcgttgccttggtaggtccagcggcggaggaactctttgatccggttcctgaacaggatctatttgaggcgctaaatgaaaccttaacgctatggaactcgccgcccgactgggctggcgatgagcgaaatgtagtgcttacgttgtcccgcatttggtacagcgcagtaaccggcaaaatcgcgccgaaggatgtcgctgccgactgggcaatggagcgcctgccggcccagtatcagcccgtcatacttgaagctagacaggcttatcttggacaagaagaagatcgcttggcctcgcgcgcagatcagttggaagaatttgtccactacgtgaaaggcgagatcaccaaggtagtcggcaaataaccctcgagccacccatgaccaaaatcccttaacgtgagttacgcgtcgttccactgagcgtcagaccccgtagaaaagatcaaaggatcttcttgagatcctttttttctgcgcgtaatctgctgcttgcaaacaaaaaaaccaccgctaccagcggtggtttgtttgccggatcaagagctaccaactctttttccgaaggtaactggcttcagcagagcgcagataccaaatactgtccttctagtgtagccgtagttaggccaccacttcaagaactctgtagcaccgcctacatacctcgctctgctaatcctgttaccagtggctgctgccagtggcgataagtcgtgtcttaccgggttggactcaagacgatagttaccggataaggcgcagcggtcgggctgaacggggggttcgtgcacacagcccagcttggagcgaacgacctacaccgaactgagatacctacagcgtgagcattgagaaagcgccacgcttcccgaagggagaaaggcggacaggtatccggtaagcggcagggtcggaacaggagagcgcacgagggagcttccagggggaaacgcctggtatctttatagtcctgtcgggtttcgccacctctgacttgagcgtcgatttttgtgatgctcgtcaggggggcggagcctatggaaaaacgccagcaacgcggcctttttacggttcctggccttttgctggccttttgctcacatgtt
Claims (21)
- 単離された非自然発生DNA結合ポリペプチドであって、
少なくとも1つのTALE反復単位;
Nキャップポリペプチド;及び
Cキャップポリペプチド
を含み、前記Cキャップポリペプチドは、TALEタンパク質のフラグメントを含む、ポリペプチド。 - 少なくとも1つのTALE反復単位は、非定型の反復可変ジ残基(repeat variable di−residue)(RVD)を含む、請求項1に記載の単離されたポリペプチド。
- 前記タンパク質は、表27に示される非定型のRVDを含む、請求項2に記載のポリペプチド。
- 前記Cキャップポリペプチドは、約230アミノ酸長未満である、請求項1〜3のいずれかに記載のポリペプチド。
- 前記Cキャップは、TALE反復ドメインを含む、請求項1〜5のいずれかに記載のポリペプチド。
- 請求項1〜5のいずれかに記載のポリペプチドと、少なくとも1つの機能ドメインと、を含む、融合タンパク質。
- 前記機能ドメインは、転写活性化因子又は転写抑制因子である、請求項6に記載の融合タンパク質。
- 前記機能ドメインは、ヌクレアーゼを含む、請求項7に記載の融合タンパク質。
- 前記ヌクレアーゼは、IIS型エンドヌクレアーゼからの少なくとも1つの切断ドメイン又は切断半ドメインを含む、請求項8に記載の融合タンパク質。
- 請求項1〜5のいずれかに記載のポリペプチド又は請求項6〜9のいずれかに記載の融合タンパク質をコードする、ポリヌクレオチド。
- 請求項1〜5のいずれかに記載のポリペプチド、請求項6〜9のいずれかに記載の融合タンパク質、又は請求項10に記載のポリヌクレオチドを含む、宿主細胞。
- 請求項1〜5のいずれかに記載のポリペプチド、請求項6〜9のいずれかに記載の融合タンパク質、又は請求項10に記載のポリヌクレオチドを含む、薬学的組成物。
- 細胞内の内在性遺伝子の発現を調節する方法であって、
前記細胞内に、請求項6〜9のいずれかに記載の融合タンパク質又は前記融合タンパク質をコードするポリヌクレオチドを導入することを含み、前記融合タンパク質は、前記内在性遺伝子内の標的部位に結合するTALE反復ドメインを含み、さらに前記内在性遺伝子の発現が調節される、方法。 - 前記調節は、遺伝子活性化を含む、請求項13に記載の方法。
- 前記調節は、遺伝子抑制又は不活性化を含む、請求項13に記載の方法。
- 前記融合タンパク質は、切断ドメイン又は切断半ドメインを含み、前記内在性遺伝子は、切断によって不活性化される、請求項15に記載の方法。
- 不活性化は、非相同末端結合(NHEJ)を介して生じる、請求項16に記載の方法。
- 細胞のゲノム内の目的とする領域を修飾する方法であって、
前記細胞に、請求項8若しくは9に記載の少なくとも1つの融合タンパク質又は前記融合タンパク質をコードするポリヌクレオチドを導入することを含み、前記融合タンパク質は、前記細胞のゲノム内で標的部位に結合するTALE反復ドメインを含み、前記融合タンパク質は、前記目的とする領域内で前記ゲノムを切断する、方法。 - 前記修飾することは、前記目的とする領域に欠失を導入することを含む、請求項18に記載の方法。
- 前記修飾することは、外因性核酸を前記目的とする領域に導入することを含み、前記方法は、前記外因性核酸を前記細胞に導入することをさらに含み、前記外因性核酸は、相同組換え又はNHEJ媒介末端捕捉によって前記目的とする領域に組み込まれる、請求項18に記載の方法。
- 前記細胞は、植物細胞、動物細胞、魚細胞、及び酵母細胞からなる群から選択される真核細胞である、請求項13〜20のいずれかに記載の方法。
Applications Claiming Priority (13)
Application Number | Priority Date | Filing Date | Title |
---|---|---|---|
US39583610P | 2010-05-17 | 2010-05-17 | |
US61/395,836 | 2010-05-17 | ||
US40142910P | 2010-08-12 | 2010-08-12 | |
US61/401,429 | 2010-08-12 | ||
US45512110P | 2010-10-13 | 2010-10-13 | |
US61/455,121 | 2010-10-13 | ||
US201061459891P | 2010-12-20 | 2010-12-20 | |
US61/459,891 | 2010-12-20 | ||
US201161462482P | 2011-02-02 | 2011-02-02 | |
US61/462,482 | 2011-02-02 | ||
US201161465869P | 2011-03-24 | 2011-03-24 | |
US61/465,869 | 2011-03-24 | ||
PCT/US2011/000885 WO2011146121A1 (en) | 2010-05-17 | 2011-05-17 | Novel dna-binding proteins and uses thereof |
Related Child Applications (1)
Application Number | Title | Priority Date | Filing Date |
---|---|---|---|
JP2016135751A Division JP2016182143A (ja) | 2010-05-17 | 2016-07-08 | 新規のdna結合タンパク質及びその使用 |
Publications (2)
Publication Number | Publication Date |
---|---|
JP2013529083A true JP2013529083A (ja) | 2013-07-18 |
JP6208580B2 JP6208580B2 (ja) | 2017-10-04 |
Family
ID=44991974
Family Applications (2)
Application Number | Title | Priority Date | Filing Date |
---|---|---|---|
JP2013511148A Active JP6208580B2 (ja) | 2010-05-17 | 2011-05-17 | 新規のdna結合タンパク質及びその使用 |
JP2016135751A Pending JP2016182143A (ja) | 2010-05-17 | 2016-07-08 | 新規のdna結合タンパク質及びその使用 |
Family Applications After (1)
Application Number | Title | Priority Date | Filing Date |
---|---|---|---|
JP2016135751A Pending JP2016182143A (ja) | 2010-05-17 | 2016-07-08 | 新規のdna結合タンパク質及びその使用 |
Country Status (9)
Country | Link |
---|---|
US (8) | US8586526B2 (ja) |
EP (2) | EP2571512B1 (ja) |
JP (2) | JP6208580B2 (ja) |
KR (1) | KR101953237B1 (ja) |
CN (1) | CN103025344B (ja) |
AU (1) | AU2011256838B2 (ja) |
CA (1) | CA2798988C (ja) |
IL (1) | IL222961B (ja) |
WO (1) | WO2011146121A1 (ja) |
Cited By (17)
Publication number | Priority date | Publication date | Assignee | Title |
---|---|---|---|---|
JP2014521317A (ja) * | 2011-07-15 | 2014-08-28 | ザ ジェネラル ホスピタル コーポレイション | 転写活性化因子様エフェクターの組立て方法 |
WO2015020218A1 (ja) * | 2013-08-09 | 2015-02-12 | 独立行政法人理化学研究所 | Dna結合タンパク質ドメインの改変による高活性taleタンパク質 |
WO2015019672A1 (ja) * | 2013-08-09 | 2015-02-12 | 国立大学法人広島大学 | Dna結合ドメインを含むポリペプチド |
WO2015068785A1 (ja) | 2013-11-06 | 2015-05-14 | 国立大学法人広島大学 | 核酸挿入用ベクター |
JP2016529904A (ja) * | 2013-08-27 | 2016-09-29 | リコンビネティクス・インコーポレイテッドRecombinetics,Inc. | 効率的な非減数分裂対立遺伝子移入 |
WO2016159110A1 (ja) * | 2015-03-31 | 2016-10-06 | 国立研究開発法人農業生物資源研究所 | バイナリー遺伝子発現システム |
JP2016175922A (ja) * | 2016-04-25 | 2016-10-06 | 国立大学法人広島大学 | Dna結合ドメインを含むポリペプチド |
US9890364B2 (en) | 2012-05-29 | 2018-02-13 | The General Hospital Corporation | TAL-Tet1 fusion proteins and methods of use thereof |
JP2018057407A (ja) * | 2013-11-18 | 2018-04-12 | クリスパー セラピューティクス アーゲー | Crispr−casシステムの材料及び方法 |
US10006011B2 (en) | 2013-08-09 | 2018-06-26 | Hiroshima University | Polypeptide containing DNA-binding domain |
US10676749B2 (en) | 2013-02-07 | 2020-06-09 | The General Hospital Corporation | Tale transcriptional activators |
JP2020110152A (ja) * | 2013-07-26 | 2020-07-27 | プレジデント アンド フェローズ オブ ハーバード カレッジ | ゲノムエンジニアリング |
WO2020209332A1 (ja) | 2019-04-09 | 2020-10-15 | 国立研究開発法人科学技術振興機構 | 核酸結合性タンパク質 |
JP2020531572A (ja) * | 2017-08-04 | 2020-11-05 | 北京大学Peking University | メチル化により修飾されたdna塩基を特異的に認識するtale rvdおよびその使用 |
US10920242B2 (en) | 2011-02-25 | 2021-02-16 | Recombinetics, Inc. | Non-meiotic allele introgression |
US11624077B2 (en) | 2017-08-08 | 2023-04-11 | Peking University | Gene knockout method |
US11891631B2 (en) | 2012-10-12 | 2024-02-06 | The General Hospital Corporation | Transcription activator-like effector (tale) - lysine-specific demethylase 1 (LSD1) fusion proteins |
Families Citing this family (553)
Publication number | Priority date | Publication date | Assignee | Title |
---|---|---|---|---|
US20120196370A1 (en) | 2010-12-03 | 2012-08-02 | Fyodor Urnov | Methods and compositions for targeted genomic deletion |
US20110239315A1 (en) | 2009-01-12 | 2011-09-29 | Ulla Bonas | Modular dna-binding domains and methods of use |
EP2206723A1 (en) | 2009-01-12 | 2010-07-14 | Bonas, Ulla | Modular DNA-binding domains |
US8772008B2 (en) * | 2009-05-18 | 2014-07-08 | Sangamo Biosciences, Inc. | Methods and compositions for increasing nuclease activity |
US20120204278A1 (en) | 2009-07-08 | 2012-08-09 | Kymab Limited | Animal models and therapeutic molecules |
US9445581B2 (en) | 2012-03-28 | 2016-09-20 | Kymab Limited | Animal models and therapeutic molecules |
DK2517557T4 (da) | 2009-07-08 | 2023-09-25 | Kymab Ltd | Dyremodeller og terapeutiske molekyler |
US8586526B2 (en) | 2010-05-17 | 2013-11-19 | Sangamo Biosciences, Inc. | DNA-binding proteins and uses thereof |
CN106834320B (zh) | 2009-12-10 | 2021-05-25 | 明尼苏达大学董事会 | Tal效应子介导的dna修饰 |
EP4328304A2 (en) | 2010-02-08 | 2024-02-28 | Sangamo Therapeutics, Inc. | Engineered cleavage half-domains |
EP2660318A1 (en) * | 2010-02-09 | 2013-11-06 | Sangamo BioSciences, Inc. | Targeted genomic modification with partially single-stranded donor molecules |
BR112012020257A8 (pt) * | 2010-02-11 | 2018-02-14 | Recombinetics Inc | métodos e aparelhos para produzir artiodátilos transgênicos |
US9567573B2 (en) | 2010-04-26 | 2017-02-14 | Sangamo Biosciences, Inc. | Genome editing of a Rosa locus using nucleases |
EP2392208B1 (en) * | 2010-06-07 | 2016-05-04 | Helmholtz Zentrum München Deutsches Forschungszentrum für Gesundheit und Umwelt (GmbH) | Fusion proteins comprising a DNA-binding domain of a Tal effector protein and a non-specific cleavage domain of a restriction nuclease and their use |
CA2805442C (en) | 2010-07-21 | 2020-05-12 | Sangamo Biosciences, Inc. | Methods and compositions for modification of an hla locus |
AU2011312562B2 (en) | 2010-09-27 | 2014-10-09 | Sangamo Therapeutics, Inc. | Methods and compositions for inhibiting viral entry into cells |
CA2814143C (en) | 2010-10-12 | 2020-09-08 | The Children's Hospital Of Philadelphia | Methods and compositions for treating hemophilia b |
US9267123B2 (en) | 2011-01-05 | 2016-02-23 | Sangamo Biosciences, Inc. | Methods and compositions for gene correction |
KR20140046408A (ko) * | 2011-02-25 | 2014-04-18 | 리컴비네틱스 인코포레이티드 | 유전자 변형 동물 및 이의 생산방법 |
SG10201602651YA (en) | 2011-04-05 | 2016-05-30 | Cellectis | Method for the generation of compact tale-nucleases and uses thereof |
WO2012145601A2 (en) | 2011-04-22 | 2012-10-26 | The Regents Of The University Of California | Adeno-associated virus virions with variant capsid and methods of use thereof |
ES2657825T3 (es) | 2011-06-06 | 2018-03-07 | Bayer Cropscience Nv | Métodos y medios para modificar el genoma de una planta en un sitio preseleccionado |
EP3613852A3 (en) | 2011-07-22 | 2020-04-22 | President and Fellows of Harvard College | Evaluation and improvement of nuclease cleavage specificity |
CA2841541C (en) | 2011-07-25 | 2019-11-12 | Sangamo Biosciences, Inc. | Methods and compositions for alteration of a cystic fibrosis transmembrane conductance regulator (cftr) gene |
US9370551B2 (en) * | 2011-07-27 | 2016-06-21 | The Broad Institute, Inc. | Compositions and methods of treating head and neck cancer |
CN103981149A (zh) | 2011-08-22 | 2014-08-13 | 拜尔作物科学公司 | 修饰植物基因组的方法和手段 |
EP3839049A3 (en) | 2011-09-19 | 2021-10-20 | Kymab Limited | Antibodies, variable domains & chains tailored for human use |
KR102209330B1 (ko) | 2011-09-21 | 2021-01-29 | 상가모 테라퓨틱스, 인코포레이티드 | 이식 유전자 발현의 조절을 위한 방법 및 조성물 |
WO2013045916A1 (en) | 2011-09-26 | 2013-04-04 | Kymab Limited | Chimaeric surrogate light chains (slc) comprising human vpreb |
US9222105B2 (en) | 2011-10-27 | 2015-12-29 | Sangamo Biosciences, Inc. | Methods and compositions for modification of the HPRT locus |
WO2013074999A1 (en) | 2011-11-16 | 2013-05-23 | Sangamo Biosciences, Inc. | Modified dna-binding proteins and uses thereof |
EP2780376B1 (en) * | 2011-11-18 | 2017-03-08 | Université Laval | Methods and products for increasing frataxin levels and uses thereof |
US8450107B1 (en) | 2011-11-30 | 2013-05-28 | The Broad Institute Inc. | Nucleotide-specific recognition sequences for designer TAL effectors |
US10801017B2 (en) | 2011-11-30 | 2020-10-13 | The Broad Institute, Inc. | Nucleotide-specific recognition sequences for designer TAL effectors |
US20130137173A1 (en) * | 2011-11-30 | 2013-05-30 | Feng Zhang | Nucleotide-specific recognition sequences for designer tal effectors |
US9253965B2 (en) | 2012-03-28 | 2016-02-09 | Kymab Limited | Animal models and therapeutic molecules |
EP3835420A1 (en) | 2011-12-05 | 2021-06-16 | Factor Bioscience Inc. | Methods and products for transfecting cells |
CA3226329A1 (en) * | 2011-12-16 | 2013-06-20 | Targetgene Biotechnologies Ltd | Compositions and methods for modifying a predetermined target nucleic acid sequence |
GB201122458D0 (en) | 2011-12-30 | 2012-02-08 | Univ Wageningen | Modified cascade ribonucleoproteins and uses thereof |
CN103193871B (zh) * | 2012-01-04 | 2018-05-15 | 清华大学 | 根据蛋白质-dna复合物晶体结构设计新型tale的方法 |
WO2013102289A1 (zh) * | 2012-01-04 | 2013-07-11 | 清华大学 | 特异结合和靶定dna-rna 杂合双链的方法 |
WO2013102290A1 (zh) * | 2012-01-04 | 2013-07-11 | 清华大学 | 特异识别含有5-甲基化胞嘧啶的dna的方法 |
CA2865011C (en) | 2012-02-29 | 2021-06-15 | Sangamo Biosciences, Inc. | Methods and compositions for treating huntington's disease |
DK3415158T3 (da) | 2012-03-15 | 2021-01-11 | Cellectis | Repeat-variable direster til targeting af nukleotider |
US9637739B2 (en) * | 2012-03-20 | 2017-05-02 | Vilnius University | RNA-directed DNA cleavage by the Cas9-crRNA complex |
US10251377B2 (en) | 2012-03-28 | 2019-04-09 | Kymab Limited | Transgenic non-human vertebrate for the expression of class-switched, fully human, antibodies |
GB2502127A (en) | 2012-05-17 | 2013-11-20 | Kymab Ltd | Multivalent antibodies and in vivo methods for their production |
US20130274129A1 (en) | 2012-04-04 | 2013-10-17 | Geneart Ag | Tal-effector assembly platform, customized services, kits and assays |
US11518997B2 (en) | 2012-04-23 | 2022-12-06 | BASF Agricultural Solutions Seed US LLC | Targeted genome engineering in plants |
CN109536526B (zh) | 2012-04-25 | 2020-06-09 | 瑞泽恩制药公司 | 核酸酶介导的使用大靶向载体的靶向 |
EP3597741A1 (en) | 2012-04-27 | 2020-01-22 | Duke University | Genetic correction of mutated genes |
US9738879B2 (en) | 2012-04-27 | 2017-08-22 | Duke University | Genetic correction of mutated genes |
CN104428414B (zh) | 2012-05-02 | 2022-01-21 | 陶氏益农公司 | 苹果酸脱氢酶的靶向修饰 |
KR102116153B1 (ko) | 2012-05-07 | 2020-05-27 | 상가모 테라퓨틱스, 인코포레이티드 | 도입유전자의 뉴클레아제-매개 표적화된 통합을 위한 방법 및 조성물 |
EA201492222A1 (ru) * | 2012-05-25 | 2015-05-29 | Селлектис | Способы конструирования неаллореактивной и устойчивой к иммуносупрессии т-клетки для иммунотерапии |
UA118014C2 (uk) | 2012-05-25 | 2018-11-12 | Те Ріджентс Оф Те Юніверсіті Оф Каліфорнія | Спосіб модифікації днк-мішені |
US20150017136A1 (en) * | 2013-07-15 | 2015-01-15 | Cellectis | Methods for engineering allogeneic and highly active t cell for immunotherapy |
US10815500B2 (en) * | 2012-06-05 | 2020-10-27 | Cellectis | Transcription activator-like effector (TALE) fusion protein |
EP2861737B1 (en) | 2012-06-19 | 2019-04-17 | Regents Of The University Of Minnesota | Gene targeting in plants using dna viruses |
AR091482A1 (es) * | 2012-06-21 | 2015-02-04 | Recombinetics Inc | Celulas modificadas geneticamente y metodos par su obtencion |
US10648001B2 (en) | 2012-07-11 | 2020-05-12 | Sangamo Therapeutics, Inc. | Method of treating mucopolysaccharidosis type I or II |
EP2872154B1 (en) | 2012-07-11 | 2017-05-31 | Sangamo BioSciences, Inc. | Methods and compositions for delivery of biologics |
EP2872625B1 (en) | 2012-07-11 | 2016-11-16 | Sangamo BioSciences, Inc. | Methods and compositions for the treatment of lysosomal storage diseases |
US20150104873A1 (en) * | 2012-07-11 | 2015-04-16 | University of Nevada, Las Vegas | Genome Surgery with Paired, Permeant Endonuclease Construct |
WO2014018601A2 (en) * | 2012-07-24 | 2014-01-30 | Cellectis | New modular base-specific nucleic acid binding domains from burkholderia rhizoxinica proteins |
CN105188767A (zh) * | 2012-07-25 | 2015-12-23 | 布罗德研究所有限公司 | 可诱导的dna结合蛋白和基因组干扰工具及其应用 |
US10058078B2 (en) | 2012-07-31 | 2018-08-28 | Recombinetics, Inc. | Production of FMDV-resistant livestock by allele substitution |
EP2890780B8 (en) | 2012-08-29 | 2020-08-19 | Sangamo Therapeutics, Inc. | Methods and compositions for treatment of a genetic condition |
KR102357105B1 (ko) * | 2012-09-04 | 2022-02-08 | 더 스크립스 리서치 인스티튜트 | 표적화된 바인딩 특이도를 갖는 키메라 폴리펩타이드들 |
UA119135C2 (uk) | 2012-09-07 | 2019-05-10 | ДАУ АГРОСАЙЄНСІЗ ЕлЕлСі | Спосіб отримання трансгенної рослини |
UA118090C2 (uk) * | 2012-09-07 | 2018-11-26 | ДАУ АГРОСАЙЄНСІЗ ЕлЕлСі | Спосіб інтегрування послідовності нуклеїнової кислоти, що представляє інтерес, у ген fad2 у клітині сої та специфічний для локусу fad2 білок, що зв'язується, здатний індукувати спрямований розрив |
CA2884162C (en) | 2012-09-07 | 2020-12-29 | Dow Agrosciences Llc | Fad3 performance loci and corresponding target site specific binding proteins capable of inducing targeted breaks |
US9597357B2 (en) | 2012-10-10 | 2017-03-21 | Sangamo Biosciences, Inc. | T cell modifying compounds and uses thereof |
CN103772506A (zh) * | 2012-10-22 | 2014-05-07 | 北京唯尚立德生物科技有限公司 | 一种“转录激活子样效应因子-功能基团-雌激素受体”功能蛋白及其应用 |
CN110669746B (zh) | 2012-10-23 | 2024-04-16 | 基因工具股份有限公司 | 用于切割靶dna的组合物及其用途 |
MX367031B (es) | 2012-11-01 | 2019-08-02 | Cellectis | Plantas para la produccion de proteinas terapeuticas. |
CA2890110C (en) | 2012-11-01 | 2023-05-02 | Factor Bioscience Inc. | Methods and products for expressing proteins in cells |
AU2013346504A1 (en) | 2012-11-16 | 2015-06-18 | Cellectis | Method for targeted modification of algae genomes |
CA2891510C (en) * | 2012-11-16 | 2022-10-18 | Transposagen Biopharmaceuticals, Inc. | Site-specific enzymes and methods of use |
CN104884626A (zh) * | 2012-11-20 | 2015-09-02 | 杰.尔.辛普洛公司 | Tal介导的转移dna插入 |
EP2929017A4 (en) | 2012-12-05 | 2016-09-28 | Sangamo Biosciences Inc | METHODS AND COMPOSITIONS FOR REGULATION OF METABOLIC DISORDERS |
IL300199A (en) * | 2012-12-06 | 2023-03-01 | Sigma Aldrich Co Llc | CRISPR-based genome modification and regulation |
SG10201912327SA (en) | 2012-12-12 | 2020-02-27 | Broad Inst Inc | Engineering and Optimization of Improved Systems, Methods and Enzyme Compositions for Sequence Manipulation |
DK2921557T3 (en) | 2012-12-12 | 2016-11-07 | Broad Inst Inc | Design of systems, methods and optimized sequence manipulation guide compositions |
EP3434776A1 (en) | 2012-12-12 | 2019-01-30 | The Broad Institute, Inc. | Methods, models, systems, and apparatus for identifying target sequences for cas enzymes or crispr-cas systems for target sequences and conveying results thereof |
US8697359B1 (en) * | 2012-12-12 | 2014-04-15 | The Broad Institute, Inc. | CRISPR-Cas systems and methods for altering expression of gene products |
DK2931898T3 (en) | 2012-12-12 | 2016-06-20 | Massachusetts Inst Technology | CONSTRUCTION AND OPTIMIZATION OF SYSTEMS, PROCEDURES AND COMPOSITIONS FOR SEQUENCE MANIPULATION WITH FUNCTIONAL DOMAINS |
PL2921557T3 (pl) * | 2012-12-12 | 2017-03-31 | Broad Inst Inc | Projektowanie systemów, sposoby i optymalizowane kompozycje kierujące do manipulacji sekwencją |
RU2721275C2 (ru) * | 2012-12-12 | 2020-05-18 | Те Брод Инститьют, Инк. | Доставка, конструирование и оптимизация систем, способов и композиций для манипуляции с последовательностями и применения в терапии |
CA2894668A1 (en) | 2012-12-12 | 2014-06-19 | The Broad Institute, Inc. | Crispr-cas systems and methods for altering expression of gene products in eukaryotic cells |
ES2741951T3 (es) * | 2012-12-17 | 2020-02-12 | Harvard College | Modificación por ingeniería genética del genoma humano guiada por ARN |
US9708589B2 (en) | 2012-12-18 | 2017-07-18 | Monsanto Technology Llc | Compositions and methods for custom site-specific DNA recombinases |
TR201808715T4 (tr) | 2012-12-21 | 2018-07-23 | Cellectis | Düşük soğuk indüklü tatlandırması olan patatesler. |
ES2904803T3 (es) | 2013-02-20 | 2022-04-06 | Regeneron Pharma | Modificación genética de ratas |
WO2014130955A1 (en) | 2013-02-25 | 2014-08-28 | Sangamo Biosciences, Inc. | Methods and compositions for enhancing nuclease-mediated gene disruption |
JP2016507244A (ja) | 2013-02-27 | 2016-03-10 | ヘルムホルツ・ツェントルム・ミュンヒェン・ドイチェス・フォルシュンクスツェントルム・フューア・ゲズントハイト・ウント・ウムベルト(ゲーエムベーハー)Helmholtz Zentrum MuenchenDeutsches Forschungszentrum fuer Gesundheit und Umwelt (GmbH) | Cas9ヌクレアーゼによる卵母細胞における遺伝子編集 |
US10240125B2 (en) | 2013-03-14 | 2019-03-26 | Immusoft Corporation | Methods for in vitro memory B cell differentiation and transduction with VSV-G pseudotyped viral vectors |
AU2014235794A1 (en) | 2013-03-14 | 2015-10-22 | Caribou Biosciences, Inc. | Compositions and methods of nucleic acid-targeting nucleic acids |
US20140363561A1 (en) * | 2013-03-15 | 2014-12-11 | J.R. Simplot Company | Tal-mediated transfer dna insertion |
US10793867B2 (en) | 2013-03-15 | 2020-10-06 | Monsanto Technology, Llc | Methods for targeted transgene-integration using custom site-specific DNA recombinases |
US11332719B2 (en) * | 2013-03-15 | 2022-05-17 | The Broad Institute, Inc. | Recombinant virus and preparations thereof |
US10113162B2 (en) | 2013-03-15 | 2018-10-30 | Cellectis | Modifying soybean oil composition through targeted knockout of the FAD2-1A/1B genes |
US11039586B2 (en) | 2013-03-15 | 2021-06-22 | Monsanto Technology Llc | Creation and transmission of megaloci |
EP2970997A1 (en) * | 2013-03-15 | 2016-01-20 | Regents of the University of Minnesota | Engineering plant genomes using crispr/cas systems |
US9788534B2 (en) | 2013-03-18 | 2017-10-17 | Kymab Limited | Animal models and therapeutic molecules |
US9828582B2 (en) * | 2013-03-19 | 2017-11-28 | Duke University | Compositions and methods for the induction and tuning of gene expression |
WO2014153470A2 (en) | 2013-03-21 | 2014-09-25 | Sangamo Biosciences, Inc. | Targeted disruption of t cell receptor genes using engineered zinc finger protein nucleases |
EP2981614A1 (en) | 2013-04-02 | 2016-02-10 | Bayer CropScience NV | Targeted genome engineering in eukaryotes |
RU2723130C2 (ru) | 2013-04-05 | 2020-06-08 | ДАУ АГРОСАЙЕНСИЗ ЭлЭлСи | Способы и композиции для встраивания экзогенной последовательности в геном растений |
WO2014169810A1 (zh) * | 2013-04-16 | 2014-10-23 | 深圳华大基因科技服务有限公司 | 分离的寡核苷酸及其用途 |
ES2699578T3 (es) | 2013-04-16 | 2019-02-11 | Regeneron Pharma | Modificación direccionada del genoma de rata |
CN103233004B (zh) * | 2013-04-28 | 2015-04-29 | 新疆农垦科学院 | 一种人工dna分子及检测目标基因表达的方法 |
US9783618B2 (en) | 2013-05-01 | 2017-10-10 | Kymab Limited | Manipulation of immunoglobulin gene diversity and multi-antibody therapeutics |
US11707056B2 (en) | 2013-05-02 | 2023-07-25 | Kymab Limited | Animals, repertoires and methods |
US9783593B2 (en) | 2013-05-02 | 2017-10-10 | Kymab Limited | Antibodies, variable domains and chains tailored for human use |
JP2016518142A (ja) | 2013-05-10 | 2016-06-23 | サンガモ バイオサイエンシーズ, インコーポレイテッド | ヌクレアーゼ媒介ゲノム遺伝子操作のための送達方法および組成物 |
US11311575B2 (en) | 2013-05-13 | 2022-04-26 | Cellectis | Methods for engineering highly active T cell for immunotherapy |
US11077144B2 (en) | 2013-05-13 | 2021-08-03 | Cellectis | CD19 specific chimeric antigen receptor and uses thereof |
US9873894B2 (en) | 2013-05-15 | 2018-01-23 | Sangamo Therapeutics, Inc. | Methods and compositions for treatment of a genetic condition |
KR20230136697A (ko) * | 2013-06-05 | 2023-09-26 | 듀크 유니버시티 | Rna-가이드 유전자 편집 및 유전자 조절 |
EP3011033B1 (en) | 2013-06-17 | 2020-02-19 | The Broad Institute, Inc. | Functional genomics using crispr-cas systems, compositions methods, screens and applications thereof |
WO2014204729A1 (en) | 2013-06-17 | 2014-12-24 | The Broad Institute Inc. | Delivery, use and therapeutic applications of the crispr-cas systems and compositions for targeting disorders and diseases using viral components |
JP6625971B2 (ja) | 2013-06-17 | 2019-12-25 | ザ・ブロード・インスティテュート・インコーポレイテッド | 配列操作のためのタンデムガイド系、方法および組成物の送達、エンジニアリングおよび最適化 |
CN113425857A (zh) | 2013-06-17 | 2021-09-24 | 布罗德研究所有限公司 | 用于肝靶向和治疗的crispr-cas系统、载体和组合物的递送与用途 |
AU2014281027A1 (en) | 2013-06-17 | 2016-01-28 | Massachusetts Institute Of Technology | Optimized CRISPR-Cas double nickase systems, methods and compositions for sequence manipulation |
FR3008424B1 (fr) * | 2013-07-09 | 2017-12-15 | Centre National De La Recherche Scient (Cnrs) | Procede de modification ciblee du genome pour la generation d'un organisme animal |
AU2014287397B2 (en) | 2013-07-10 | 2019-10-10 | President And Fellows Of Harvard College | Orthogonal Cas9 proteins for RNA-guided gene regulation and editing |
US20160264982A1 (en) * | 2013-07-16 | 2016-09-15 | Shanghai Institutes For Biological Sciences, Chinese Academy Of Sciences | Method for plant genome site-directed modification |
WO2015017866A1 (en) | 2013-08-02 | 2015-02-05 | Enevolv, Inc. | Processes and host cells for genome, pathway, and biomolecular engineering |
US9163284B2 (en) | 2013-08-09 | 2015-10-20 | President And Fellows Of Harvard College | Methods for identifying a target site of a Cas9 nuclease |
US9359599B2 (en) | 2013-08-22 | 2016-06-07 | President And Fellows Of Harvard College | Engineered transcription activator-like effector (TALE) domains and uses thereof |
EP3988654A1 (en) | 2013-08-28 | 2022-04-27 | Sangamo Therapeutics, Inc. | Compositions for linking dna-binding domains and cleavage domains |
WO2015033343A1 (en) | 2013-09-03 | 2015-03-12 | Yissum Research Development Company Of The Hebrew University Of Jerusalem Ltd. | Compositions and methods for expressing recombinant polypeptides |
US9228207B2 (en) | 2013-09-06 | 2016-01-05 | President And Fellows Of Harvard College | Switchable gRNAs comprising aptamers |
US9526784B2 (en) | 2013-09-06 | 2016-12-27 | President And Fellows Of Harvard College | Delivery system for functional nucleases |
US20150079680A1 (en) | 2013-09-18 | 2015-03-19 | Kymab Limited | Methods, cells & organisms |
EP3051942B1 (en) | 2013-10-01 | 2020-09-02 | Kymab Limited | Animal models and therapeutic molecules |
US9476884B2 (en) | 2013-10-04 | 2016-10-25 | University Of Massachusetts | Hybridization- independent labeling of repetitive DNA sequence in human chromosomes |
EP3057432B1 (en) | 2013-10-17 | 2018-11-21 | Sangamo Therapeutics, Inc. | Delivery methods and compositions for nuclease-mediated genome engineering in hematopoietic stem cells |
CN110713995B (zh) | 2013-10-17 | 2023-08-01 | 桑格摩生物科学股份有限公司 | 用于核酸酶介导的基因组工程改造的递送方法和组合物 |
JP6484987B2 (ja) | 2013-10-21 | 2019-03-20 | Jnc株式会社 | エビルシフェラーゼの触媒蛋白質の変異遺伝子とその使用法 |
WO2015059690A1 (en) | 2013-10-24 | 2015-04-30 | Yeda Research And Development Co. Ltd. | Polynucleotides encoding brex system polypeptides and methods of using s ame |
MX2016005379A (es) | 2013-10-25 | 2017-02-15 | Livestock Improvement Corp Ltd | Marcadores geneticos y usos de los mismos. |
WO2015065964A1 (en) | 2013-10-28 | 2015-05-07 | The Broad Institute Inc. | Functional genomics using crispr-cas systems, compositions, methods, screens and applications thereof |
UA121459C2 (uk) | 2013-11-04 | 2020-06-10 | Дау Агросаєнсиз Елелсі | Рекомбінантна молекула нуклеїнової кислоти для трансформації рослини сої |
BR102014027436B1 (pt) | 2013-11-04 | 2022-06-28 | Dow Agrosciences Llc | Molécula de ácido nucleico recombinante, e método para a produção de uma célula vegetal transgênica |
EP3066110B1 (en) | 2013-11-04 | 2021-12-29 | Corteva Agriscience LLC | Optimal maize loci |
JP6633532B2 (ja) | 2013-11-04 | 2020-01-22 | ダウ アグロサイエンシィズ エルエルシー | 最適なトウモロコシ遺伝子座 |
CN111218447A (zh) | 2013-11-07 | 2020-06-02 | 爱迪塔斯医药有限公司 | 使用统治型gRNA的CRISPR相关方法和组合物 |
EP3068905A4 (en) | 2013-11-11 | 2017-07-05 | Sangamo BioSciences, Inc. | Methods and compositions for treating huntington's disease |
PT3492593T (pt) | 2013-11-13 | 2021-10-18 | Childrens Medical Center | Regulação da expressão de genes mediada por nucleases |
US9951353B2 (en) * | 2013-11-15 | 2018-04-24 | The United States Of America, As Represented By The Secretary, Dept. Of Health And Human Services | Engineering neural stem cells using homologous recombination |
EP2878667A1 (en) * | 2013-11-29 | 2015-06-03 | Institut Pasteur | TAL effector means useful for partial or full deletion of DNA tandem repeats |
CN105940109A (zh) | 2013-12-06 | 2016-09-14 | 国立健康与医学研究所 | 用于在受试者的视网膜色素上皮中表达目的多核苷酸的方法和药物组合物 |
CN105940013B (zh) | 2013-12-09 | 2020-03-27 | 桑格摩生物科学股份有限公司 | 用于治疗血友病的方法和组合物 |
KR102170502B1 (ko) | 2013-12-11 | 2020-10-28 | 리제너론 파마슈티칼스 인코포레이티드 | 게놈의 표적화된 변형을 위한 방법 및 조성물 |
US20150166985A1 (en) | 2013-12-12 | 2015-06-18 | President And Fellows Of Harvard College | Methods for correcting von willebrand factor point mutations |
KR20160097327A (ko) | 2013-12-12 | 2016-08-17 | 더 브로드 인스티튜트, 인코퍼레이티드 | 유전자 산물, 구조 정보 및 유도성 모듈형 cas 효소의 발현의 변경을 위한 crispr-cas 시스템 및 방법 |
CA2932436A1 (en) | 2013-12-12 | 2015-06-18 | The Broad Institute, Inc. | Compositions and methods of use of crispr-cas systems in nucleotide repeat disorders |
WO2015089364A1 (en) | 2013-12-12 | 2015-06-18 | The Broad Institute Inc. | Crystal structure of a crispr-cas system, and uses thereof |
JP6793547B2 (ja) | 2013-12-12 | 2020-12-02 | ザ・ブロード・インスティテュート・インコーポレイテッド | 最適化機能CRISPR−Cas系による配列操作のための系、方法および組成物 |
SG10201804976YA (en) | 2013-12-12 | 2018-07-30 | Broad Inst Inc | Delivery, Use and Therapeutic Applications of the Crispr-Cas Systems and Compositions for Genome Editing |
AU2014368892B2 (en) | 2013-12-20 | 2019-06-13 | Fred Hutchinson Cancer Center | Tagged chimeric effector molecules and receptors thereof |
US10774338B2 (en) | 2014-01-16 | 2020-09-15 | The Regents Of The University Of California | Generation of heritable chimeric plant traits |
KR102415811B1 (ko) | 2014-01-31 | 2022-07-04 | 팩터 바이오사이언스 인크. | 핵산 제조 및 송달을 위한 방법 및 산물 |
SI3102673T1 (sl) | 2014-02-03 | 2020-11-30 | Sangamo Therapeutics, Inc. | Postopki in sestavki za zdravljenje beta talasemije |
CN113215219A (zh) | 2014-02-13 | 2021-08-06 | 宝生物工程(美国) 有限公司 | 从核酸的初始集合中耗尽靶分子的方法、以及用于实践其的组合物和试剂盒 |
CN104844696A (zh) * | 2014-02-19 | 2015-08-19 | 北京大学 | 一种转录激活子样效应因子功能蛋白设计、合成及其应用 |
WO2015127439A1 (en) | 2014-02-24 | 2015-08-27 | Sangamo Biosciences, Inc. | Methods and compositions for nuclease-mediated targeted integration |
TW201538518A (zh) | 2014-02-28 | 2015-10-16 | Dow Agrosciences Llc | 藉由嵌合基因調控元件所賦予之根部特異性表現 |
EP3114227B1 (en) | 2014-03-05 | 2021-07-21 | Editas Medicine, Inc. | Crispr/cas-related methods and compositions for treating usher syndrome and retinitis pigmentosa |
US11141493B2 (en) | 2014-03-10 | 2021-10-12 | Editas Medicine, Inc. | Compositions and methods for treating CEP290-associated disease |
EP3116997B1 (en) | 2014-03-10 | 2019-05-15 | Editas Medicine, Inc. | Crispr/cas-related methods and compositions for treating leber's congenital amaurosis 10 (lca10) |
US11339437B2 (en) | 2014-03-10 | 2022-05-24 | Editas Medicine, Inc. | Compositions and methods for treating CEP290-associated disease |
EP3929279A1 (en) | 2014-03-18 | 2021-12-29 | Sangamo Therapeutics, Inc. | Methods and compositions for regulation of zinc finger protein expression |
EP3981876A1 (en) | 2014-03-26 | 2022-04-13 | Editas Medicine, Inc. | Crispr/cas-related methods and compositions for treating sickle cell disease |
US10507232B2 (en) | 2014-04-02 | 2019-12-17 | University Of Florida Research Foundation, Incorporated | Materials and methods for the treatment of latent viral infection |
CN103923215B (zh) * | 2014-04-15 | 2016-02-10 | 中国农业大学 | 使ACC-α基因启动子PⅢ失活的物质及其应用 |
JP6265020B2 (ja) * | 2014-04-16 | 2018-01-24 | Jnc株式会社 | エビルシフェラーゼの触媒蛋白質の変異遺伝子とその使用法 |
US9522936B2 (en) | 2014-04-24 | 2016-12-20 | Sangamo Biosciences, Inc. | Engineered transcription activator like effector (TALE) proteins |
DE102014106327A1 (de) | 2014-05-07 | 2015-11-12 | Universitätsklinikum Hamburg-Eppendorf (UKE) | TAL-Effektornuklease zum gezielten Knockout des HIV-Korezeptors CCR5 |
CN106413760B (zh) | 2014-05-08 | 2020-01-14 | 桑格摩生物科学股份有限公司 | 用于治疗亨廷顿病的方法和组合物 |
WO2015175642A2 (en) | 2014-05-13 | 2015-11-19 | Sangamo Biosciences, Inc. | Methods and compositions for prevention or treatment of a disease |
EP3152319A4 (en) | 2014-06-05 | 2017-12-27 | Sangamo BioSciences, Inc. | Methods and compositions for nuclease design |
CA2952906A1 (en) | 2014-06-20 | 2015-12-23 | Cellectis | Potatoes with reduced granule-bound starch synthase |
CA2954201A1 (en) | 2014-07-08 | 2016-01-14 | Vib Vzw | Means and methods to increase plant yield |
WO2016005985A2 (en) | 2014-07-09 | 2016-01-14 | Yissum Research Development Company Of The Hebrew University Of Jerusalem Ltd. | Method for reprogramming cells |
MX2017000555A (es) | 2014-07-14 | 2017-08-10 | Univ Washington State | Desactivacion de nanos que produce ablacion de celulas de linea germinal. |
US10738278B2 (en) | 2014-07-15 | 2020-08-11 | Juno Therapeutics, Inc. | Engineered cells for adoptive cell therapy |
WO2016014837A1 (en) | 2014-07-25 | 2016-01-28 | Sangamo Biosciences, Inc. | Gene editing for hiv gene therapy |
US9816074B2 (en) | 2014-07-25 | 2017-11-14 | Sangamo Therapeutics, Inc. | Methods and compositions for modulating nuclease-mediated genome engineering in hematopoietic stem cells |
CA2956224A1 (en) | 2014-07-30 | 2016-02-11 | President And Fellows Of Harvard College | Cas9 proteins including ligand-dependent inteins |
WO2016019144A2 (en) | 2014-07-30 | 2016-02-04 | Sangamo Biosciences, Inc. | Gene correction of scid-related genes in hematopoietic stem and progenitor cells |
ES2746529T3 (es) | 2014-09-04 | 2020-03-06 | Memorial Sloan Kettering Cancer Center | Terapia génica de globina para tratar hemoglobinopatías |
IL234638A0 (en) | 2014-09-14 | 2014-12-02 | Yeda Res & Dev | NMDA receptor antagonists for the treatment of Gaucher disease |
CN113699113A (zh) | 2014-09-16 | 2021-11-26 | 桑格摩治疗股份有限公司 | 用于造血干细胞中核酸酶介导的基因组工程化和校正的方法和组合物 |
CN105440111B (zh) * | 2014-09-30 | 2019-08-13 | 深圳华大基因研究院 | 一对转录激活子样效应因子核酸酶及其编码序列与应用 |
WO2016054326A1 (en) | 2014-10-01 | 2016-04-07 | The General Hospital Corporation | Methods for increasing efficiency of nuclease-induced homology-directed repair |
WO2016077273A1 (en) * | 2014-11-11 | 2016-05-19 | Q Therapeutics, Inc. | Engineering mesenchymal stem cells using homologous recombination |
WO2016081924A1 (en) | 2014-11-20 | 2016-05-26 | Duke University | Compositions, systems and methods for cell therapy |
US10457960B2 (en) | 2014-11-21 | 2019-10-29 | Regeneron Pharmaceuticals, Inc. | Methods and compositions for targeted genetic modification using paired guide RNAs |
WO2016094867A1 (en) | 2014-12-12 | 2016-06-16 | The Broad Institute Inc. | Protected guide rnas (pgrnas) |
US10889834B2 (en) | 2014-12-15 | 2021-01-12 | Sangamo Therapeutics, Inc. | Methods and compositions for enhancing targeted transgene integration |
CN107109427B (zh) | 2014-12-23 | 2021-06-18 | 先正达参股股份有限公司 | 用于鉴定和富集包含位点特异性基因组修饰的细胞的方法和组合物 |
EP3247366A4 (en) | 2015-01-21 | 2018-10-31 | Sangamo Therapeutics, Inc. | Methods and compositions for identification of highly specific nucleases |
JP6929791B2 (ja) | 2015-02-09 | 2021-09-01 | デューク ユニバーシティ | エピゲノム編集のための組成物および方法 |
EP3256585A4 (en) | 2015-02-13 | 2018-08-15 | Factor Bioscience Inc. | Nucleic acid products and methods of administration thereof |
WO2016135507A1 (en) * | 2015-02-27 | 2016-09-01 | University Of Edinburgh | Nucleic acid editing systems |
CN104694539A (zh) * | 2015-03-09 | 2015-06-10 | 上海市同济医院 | 一种利用talen技术构建可敲除小鼠cramp基因质粒的方法 |
EP3273971A4 (en) | 2015-03-27 | 2018-12-05 | Yeda Research and Development Co. Ltd. | Methods of treating motor neuron diseases |
EP4335918A3 (en) | 2015-04-03 | 2024-04-17 | Dana-Farber Cancer Institute, Inc. | Composition and methods of genome editing of b-cells |
JP2018511333A (ja) | 2015-04-15 | 2018-04-26 | ダウ アグロサイエンシィズ エルエルシー | 導入遺伝子発現のための植物プロモータ |
KR20170136573A (ko) | 2015-04-15 | 2017-12-11 | 다우 아그로사이언시즈 엘엘씨 | 트랜스젠 발현을 위한 식물 프로모터 |
KR102535217B1 (ko) | 2015-04-24 | 2023-05-19 | 에디타스 메디신, 인코포레이티드 | Cas9 분자/가이드 rna 분자 복합체의 평가 |
US11827904B2 (en) | 2015-04-29 | 2023-11-28 | Fred Hutchinson Cancer Center | Modified stem cells and uses thereof |
US10179918B2 (en) | 2015-05-07 | 2019-01-15 | Sangamo Therapeutics, Inc. | Methods and compositions for increasing transgene activity |
MX2017014446A (es) | 2015-05-12 | 2018-06-13 | Sangamo Therapeutics Inc | Regulacion de expresion genica mediada por nucleasa. |
EP3303586A1 (en) | 2015-05-29 | 2018-04-11 | Juno Therapeutics, Inc. | Composition and methods for regulating inhibitory interactions in genetically engineered cells |
ES2886599T3 (es) | 2015-06-17 | 2021-12-20 | Poseida Therapeutics Inc | Composiciones y métodos para dirigir proteínas a loci específicos en el genoma |
WO2016205759A1 (en) | 2015-06-18 | 2016-12-22 | The Broad Institute Inc. | Engineering and optimization of systems, methods, enzymes and guide scaffolds of cas9 orthologs and variants for sequence manipulation |
US9957501B2 (en) | 2015-06-18 | 2018-05-01 | Sangamo Therapeutics, Inc. | Nuclease-mediated regulation of gene expression |
RU2021120582A (ru) | 2015-06-18 | 2021-09-02 | Те Брод Инститьют, Инк. | Мутации фермента crispr, уменьшающие нецелевые эффекты |
ES2594486B1 (es) * | 2015-06-19 | 2017-09-26 | Biopraxis Research Aie | Molécula de ácido nucleico, proteína de fusión y método para modificar el material genético de una célula |
ES2895652T3 (es) | 2015-07-07 | 2022-02-22 | Inst Nat Sante Rech Med | Métodos y composiciones farmacéuticas para expresar un polinucleótido de interés en el sistema nervioso periférico de un sujeto |
CA2991301A1 (en) | 2015-07-13 | 2017-01-19 | Sangamo Therapeutics, Inc. | Delivery methods and compositions for nuclease-mediated genome engineering |
MA42895A (fr) | 2015-07-15 | 2018-05-23 | Juno Therapeutics Inc | Cellules modifiées pour thérapie cellulaire adoptive |
MX2017016844A (es) | 2015-07-16 | 2018-08-15 | Biokine Therapeutics Ltd | Composiciones y procedimientos para tratar cancer. |
FI3331355T3 (fi) | 2015-08-06 | 2024-06-17 | Univ Missouri | Sikojen lisääntymishäiriö- ja keuhkotulehdusoireyhtymävirukselle (PRRSV) resistentti sika ja soluja, joilla on muokattuja CD163-geenejä |
ES2929110T3 (es) | 2015-08-25 | 2022-11-24 | Univ Duke | Composiciones y métodos para mejorar la especificidad en ingeniería genética usando endonucleasas guiadas por ARN |
US10837024B2 (en) | 2015-09-17 | 2020-11-17 | Cellectis | Modifying messenger RNA stability in plant transformations |
TW201718861A (zh) | 2015-09-22 | 2017-06-01 | 道禮責任有限公司 | 用於轉殖基因表現之植物啟動子及3’utr |
TW201718862A (zh) | 2015-09-22 | 2017-06-01 | Dow Agrosciences Llc | 用於轉殖基因表現之植物啟動子及3’utr |
IL258030B (en) | 2015-09-23 | 2022-08-01 | Sangamo Therapeutics Inc | Htt suppressing factors and their uses |
US11970710B2 (en) | 2015-10-13 | 2024-04-30 | Duke University | Genome engineering with Type I CRISPR systems in eukaryotic cells |
US10280429B2 (en) | 2015-10-22 | 2019-05-07 | Dow Agrosciences Llc | Plant promoter for transgene expression |
US20190225955A1 (en) | 2015-10-23 | 2019-07-25 | President And Fellows Of Harvard College | Evolved cas9 proteins for gene editing |
AU2016343887B2 (en) | 2015-10-28 | 2023-04-06 | Sangamo Therapeutics, Inc. | Liver-specific constructs, factor VIII expression cassettes and methods of use thereof |
WO2017079428A1 (en) | 2015-11-04 | 2017-05-11 | President And Fellows Of Harvard College | Site specific germline modification |
WO2017078935A1 (en) | 2015-11-04 | 2017-05-11 | Dow Agrosciences Llc | Plant promoter for transgene expression |
EP3380622A4 (en) | 2015-11-23 | 2019-08-07 | Sangamo Therapeutics, Inc. | METHODS AND COMPOSITIONS FOR MODIFYING IMMUNITY |
CA3008413A1 (en) | 2015-12-18 | 2017-06-22 | Sangamo Therapeutics, Inc. | Targeted disruption of the t cell receptor |
BR112018012235A2 (pt) | 2015-12-18 | 2018-12-04 | Sangamo Therapeutics Inc | rompimento alvejado do receptor de células de mhc |
JP2019505511A (ja) | 2016-01-06 | 2019-02-28 | イェダ リサーチ アンド ディベロップメント カンパニー リミテッドYeda Research And Development Co.Ltd. | 悪性疾患、自己免疫疾患および炎症性疾患を処置するための組成物および方法 |
US20170202931A1 (en) | 2016-01-15 | 2017-07-20 | Sangamo Therapeutics, Inc. | Methods and compositions for the treatment of neurologic disease |
WO2017125931A1 (en) | 2016-01-21 | 2017-07-27 | The State Of Israel, Ministry Of Agriculture & Rural Development, Agricultural Research Organization (Aro) (Volcani Center) | Parthenocarpic plants and methods of producing same |
WO2017130205A1 (en) | 2016-01-31 | 2017-08-03 | Hadasit Medical Research Services And Development Ltd. | Autosomal-identical pluripotent stem cell populations having non-identical sex chromosomal composition and uses thereof |
JP7012650B2 (ja) | 2016-02-02 | 2022-01-28 | サンガモ セラピューティクス, インコーポレイテッド | Dna結合ドメインと切断ドメインとを連結するための組成物 |
EP3410843A1 (en) | 2016-02-02 | 2018-12-12 | Cellectis | Modifying soybean oil composition through targeted knockout of the fad3a/b/c genes |
WO2017135316A1 (ja) * | 2016-02-04 | 2017-08-10 | 花王株式会社 | 変異リゾプス属菌 |
WO2017138008A2 (en) | 2016-02-14 | 2017-08-17 | Yeda Research And Development Co. Ltd. | Methods of modulating protein exocytosis and uses of same in therapy |
US20190216891A1 (en) | 2016-03-06 | 2019-07-18 | Yeda Research And Development Co., Ltd. | Method for modulating myelination |
EP3433362B1 (en) | 2016-03-23 | 2021-05-05 | Dana-Farber Cancer Institute, Inc. | Methods for enhancing the efficiency of gene editing |
EP3433364A1 (en) | 2016-03-25 | 2019-01-30 | Editas Medicine, Inc. | Systems and methods for treating alpha 1-antitrypsin (a1at) deficiency |
CA3022600A1 (en) | 2016-04-29 | 2017-11-02 | Adverum Biotechnologies, Inc. | Evasion of neutralizing antibodies by a recombinant adeno-associated virus |
CR20180589A (es) | 2016-05-13 | 2019-02-27 | 4D Molecular Therapeutics Inc | Cápsulas variantes de virus adenoasociados y métodos de uso de estas |
CN109476715B (zh) | 2016-05-26 | 2023-06-27 | 纽海姆有限公司 | 产无籽果实的植物 |
AU2017268842B2 (en) | 2016-05-27 | 2022-08-11 | Aadigen, Llc | Peptides and nanoparticles for intracellular delivery of genome-editing molecules |
GB2552861B (en) | 2016-06-02 | 2019-05-15 | Sigma Aldrich Co Llc | Using programmable DNA binding proteins to enhance targeted genome modification |
MA45491A (fr) | 2016-06-27 | 2019-05-01 | Juno Therapeutics Inc | Épitopes à restriction cmh-e, molécules de liaison et procédés et utilisations associés |
MA45455A (fr) | 2016-06-27 | 2019-05-01 | Juno Therapeutics Inc | Procédé d'identification d'épitopes peptidiques, molécules qui se lient à de tels épitopes et utilisations associées |
US11344511B2 (en) | 2016-07-27 | 2022-05-31 | Case Western Reserve University | Compounds and methods of promoting myelination |
EP3827812A1 (en) | 2016-07-29 | 2021-06-02 | The Regents of the University of California | Adeno-associated virus virions with variant capsid and methods of use thereof |
CA3032822A1 (en) | 2016-08-02 | 2018-02-08 | Editas Medicine, Inc. | Compositions and methods for treating cep290 associated disease |
WO2018027078A1 (en) | 2016-08-03 | 2018-02-08 | President And Fellows Of Harard College | Adenosine nucleobase editors and uses thereof |
US11193136B2 (en) | 2016-08-09 | 2021-12-07 | Vib Vzw | Cellulose synthase inhibitors and mutant plants |
AU2017308889B2 (en) | 2016-08-09 | 2023-11-09 | President And Fellows Of Harvard College | Programmable Cas9-recombinase fusion proteins and uses thereof |
CA2941315C (en) | 2016-08-12 | 2018-03-06 | Api Labs Inc. | High thebaine poppy and methods of producing the same |
CA3033374A1 (en) | 2016-08-15 | 2018-02-22 | Enevolv, Inc. | Molecule sensor systems |
US10576167B2 (en) | 2016-08-17 | 2020-03-03 | Factor Bioscience Inc. | Nucleic acid products and methods of administration thereof |
WO2018035354A1 (en) | 2016-08-17 | 2018-02-22 | Monsanto Technology Llc | Methods and compositions for short stature plants through manipulation of gibberellin metabolism to increase harvestable yield |
IL247368A0 (en) | 2016-08-18 | 2016-11-30 | Yeda Res & Dev | Diagnostic and therapeutic uses of exosomes |
SG10201913948PA (en) | 2016-08-24 | 2020-03-30 | Sangamo Therapeutics Inc | Engineered target specific nucleases |
US11542509B2 (en) | 2016-08-24 | 2023-01-03 | President And Fellows Of Harvard College | Incorporation of unnatural amino acids into proteins using base editing |
PE20190568A1 (es) | 2016-08-24 | 2019-04-22 | Sangamo Therapeutics Inc | Regulacion de la expresion genica usando nucleasas modificadas |
US10960085B2 (en) | 2016-09-07 | 2021-03-30 | Sangamo Therapeutics, Inc. | Modulation of liver genes |
JP6866475B2 (ja) | 2016-09-23 | 2021-04-28 | フレッド ハッチンソン キャンサー リサーチ センター | マイナー組織適合性(h)抗原ha−1に特異的なtcr及びその使用 |
KR20190062505A (ko) | 2016-10-03 | 2019-06-05 | 주노 쎄러퓨티크스 인코퍼레이티드 | Hpv-특이적 결합 분자 |
CN110291199B (zh) | 2016-10-03 | 2023-12-22 | 美国陶氏益农公司 | 用于转基因表达的植物启动子 |
US10519459B2 (en) | 2016-10-03 | 2019-12-31 | Dow Agrosciences Llc | Plant promoter from Panicum virgatum |
EP3523422A1 (en) | 2016-10-05 | 2019-08-14 | FUJIFILM Cellular Dynamics, Inc. | Generating mature lineages from induced pluripotent stem cells with mecp2 disruption |
BR112019006652A2 (pt) | 2016-10-13 | 2019-07-02 | Juno Therapeutics Inc | métodos e composições para imunoterapia envolvendo moduladores da via metabólica do triptofano |
KR20240007715A (ko) | 2016-10-14 | 2024-01-16 | 프레지던트 앤드 펠로우즈 오브 하바드 칼리지 | 핵염기 에디터의 aav 전달 |
GB201617559D0 (en) | 2016-10-17 | 2016-11-30 | University Court Of The University Of Edinburgh The | Swine comprising modified cd163 and associated methods |
US11192925B2 (en) | 2016-10-19 | 2021-12-07 | Adverum Biotechnologies, Inc. | Modified AAV capsids and uses thereof |
WO2018075736A1 (en) | 2016-10-20 | 2018-04-26 | Sangamo Therapeutics, Inc. | Methods and compositions for the treatment of fabry disease |
CA3041668A1 (en) | 2016-10-31 | 2018-05-03 | Sangamo Therapeutics, Inc. | Gene correction of scid-related genes in hematopoietic stem and progenitor cells |
US11312972B2 (en) | 2016-11-16 | 2022-04-26 | Cellectis | Methods for altering amino acid content in plants through frameshift mutations |
US11542508B2 (en) | 2016-11-28 | 2023-01-03 | Yeda Research And Development Co. Ltd. | Isolated polynucleotides and polypeptides and methods of using same for expressing an expression product of interest |
WO2018102665A1 (en) | 2016-12-01 | 2018-06-07 | Sangamo Therapeutics, Inc. | Tau modulators and methods and compositions for delivery thereof |
WO2018102612A1 (en) | 2016-12-02 | 2018-06-07 | Juno Therapeutics, Inc. | Engineered b cells and related compositions and methods |
CA3045338A1 (en) | 2016-12-05 | 2018-06-14 | Juno Therapeutics, Inc. | Production of engineered cells for adoptive cell therapy |
JP7416406B2 (ja) | 2016-12-08 | 2024-01-17 | ケース ウェスタン リザーブ ユニバーシティ | 機能的ミエリン産生を増進させるための方法および組成物 |
US20190323019A1 (en) * | 2016-12-21 | 2019-10-24 | Yi-Hua Hsu | Composition for editing a nucleic acid sequence and method using the same |
WO2018119359A1 (en) | 2016-12-23 | 2018-06-28 | President And Fellows Of Harvard College | Editing of ccr5 receptor gene to protect against hiv infection |
EP3563159A1 (en) | 2016-12-29 | 2019-11-06 | Ukko Inc. | Methods for identifying and de-epitoping allergenic polypeptides |
WO2018129021A1 (en) * | 2017-01-06 | 2018-07-12 | The Johns Hopkins University | Modular, inducible repressors for the control of gene expression |
IL250479A0 (en) | 2017-02-06 | 2017-03-30 | Sorek Rotem | Genetically engineered cells expressing a disarm system that confers resistance to phages and methods for their production |
EP3592853A1 (en) | 2017-03-09 | 2020-01-15 | President and Fellows of Harvard College | Suppression of pain by gene editing |
WO2018165629A1 (en) | 2017-03-10 | 2018-09-13 | President And Fellows Of Harvard College | Cytosine to guanine base editor |
EP3596217A1 (en) | 2017-03-14 | 2020-01-22 | Editas Medicine, Inc. | Systems and methods for the treatment of hemoglobinopathies |
WO2018170338A2 (en) | 2017-03-15 | 2018-09-20 | Fred Hutchinson Cancer Research Center | High affinity mage-a1-specific tcrs and uses thereof |
IL269458B2 (en) | 2017-03-23 | 2024-02-01 | Harvard College | Nucleic base editors that include nucleic acid programmable DNA binding proteins |
AU2018251150B2 (en) | 2017-04-13 | 2024-05-09 | Albert-Ludwigs-Universität Freiburg | New sequence specific reagents targeting CCR5 in primary hematopoietic cells |
EP3615668B1 (en) | 2017-04-25 | 2024-02-28 | Cellectis | Alfalfa with reduced lignin composition |
CA3061612A1 (en) | 2017-04-28 | 2018-11-01 | Acuitas Therapeutics, Inc. | Novel carbonyl lipids and lipid nanoparticle formulations for delivery of nucleic acids |
RU2019139045A (ru) | 2017-05-03 | 2021-06-03 | Сангамо Терапьютикс, Инк. | Способы и композиции для модификации гена регулятора трансмембранной проводимости при кистозном фиброзе (cftr) |
IL252151A0 (en) | 2017-05-07 | 2017-07-31 | Fainzilber Michael | Treatment of stress disorders |
EP3622070A2 (en) | 2017-05-10 | 2020-03-18 | Editas Medicine, Inc. | Crispr/rna-guided nuclease systems and methods |
WO2018209320A1 (en) | 2017-05-12 | 2018-11-15 | President And Fellows Of Harvard College | Aptazyme-embedded guide rnas for use with crispr-cas9 in genome editing and transcriptional activation |
US10780119B2 (en) | 2017-05-24 | 2020-09-22 | Effector Therapeutics Inc. | Methods and compositions for cellular immunotherapy |
US11512287B2 (en) | 2017-06-16 | 2022-11-29 | Sangamo Therapeutics, Inc. | Targeted disruption of T cell and/or HLA receptors |
ES2871400T3 (es) | 2017-06-20 | 2021-10-28 | Inst Nat Sante Rech Med | Células inmunitarias deficientes en SUV39H1 |
IL253642A0 (en) | 2017-07-24 | 2017-09-28 | Seger Rony | Combined treatment for cancer |
US11732274B2 (en) | 2017-07-28 | 2023-08-22 | President And Fellows Of Harvard College | Methods and compositions for evolving base editors using phage-assisted continuous evolution (PACE) |
BR112020001364A2 (pt) | 2017-07-31 | 2020-08-11 | Regeneron Pharmaceuticals, Inc. | métodos para testar e modificar a capacidade de uma crispr/cas nuclease. |
RU2019143568A (ru) | 2017-07-31 | 2021-09-02 | Регенерон Фармасьютикалс, Инк. | Способы и композиции для оценки crispr/cas-индуцированной рекомбинации с экзогенной донорной нуклеиновой кислотой in vivo |
US11021719B2 (en) | 2017-07-31 | 2021-06-01 | Regeneron Pharmaceuticals, Inc. | Methods and compositions for assessing CRISPER/Cas-mediated disruption or excision and CRISPR/Cas-induced recombination with an exogenous donor nucleic acid in vivo |
CN109384833B (zh) * | 2017-08-04 | 2021-04-27 | 北京大学 | 特异性识别甲基化修饰dna碱基的tale rvd及其应用 |
US11692166B2 (en) | 2017-08-23 | 2023-07-04 | Technion Research & Development Foundation Limited | Compositions and methods for improving alcohol tolerance in yeast |
US11319532B2 (en) | 2017-08-30 | 2022-05-03 | President And Fellows Of Harvard College | High efficiency base editors comprising Gam |
JP7407701B2 (ja) | 2017-09-06 | 2024-01-04 | フレッド ハッチンソン キャンサー センター | strepタグ特異的キメラ受容体およびその使用 |
US20210106618A1 (en) | 2017-09-06 | 2021-04-15 | Fred Hutchinson Cancer Research Center | Methods for improving adoptive cell therapy |
CR20200165A (es) | 2017-09-20 | 2020-06-05 | 4D Molecular Therapeutics Inc | Variantes de cápsides de virus adenoasociados y métodos de uso de estas |
MX2020003536A (es) | 2017-10-03 | 2020-09-14 | Juno Therapeutics Inc | Moleculas de union especifica a virus de papiloma humano (hpv). |
EP3697906A1 (en) | 2017-10-16 | 2020-08-26 | The Broad Institute, Inc. | Uses of adenosine base editors |
EP3697435A1 (en) | 2017-10-20 | 2020-08-26 | Fred Hutchinson Cancer Research Center | Compositions and methods of immunotherapy targeting tigit and/or cd112r or comprising cd226 overexpression |
US11851679B2 (en) | 2017-11-01 | 2023-12-26 | Juno Therapeutics, Inc. | Method of assessing activity of recombinant antigen receptors |
IL255664A0 (en) | 2017-11-14 | 2017-12-31 | Shachar Idit | Hematopoietic stem cells with enhanced properties |
AU2018368786A1 (en) | 2017-11-17 | 2020-06-18 | Iovance Biotherapeutics, Inc. | TIL expansion from fine needle aspirates and small biopsies |
JP2021503885A (ja) | 2017-11-22 | 2021-02-15 | アイオバンス バイオセラピューティクス,インコーポレイテッド | 末梢血からの末梢血リンパ球(pbl)の拡大培養 |
CN111770999A (zh) | 2017-11-27 | 2020-10-13 | 4D分子治疗有限公司 | 腺相关病毒变体衣壳和用于抑制血管生成的应用 |
US20220267403A1 (en) | 2017-12-01 | 2022-08-25 | Fred Hutchinson Cancer Research Center | Binding proteins specific for 5t4 and uses thereof |
WO2019113321A1 (en) | 2017-12-06 | 2019-06-13 | Memorial Sloan-Kettering Cancer Center | Globin gene therapy for treating hemoglobinopathies |
EP3501268B1 (en) | 2017-12-22 | 2021-09-15 | KWS SAAT SE & Co. KGaA | Regeneration of plants in the presence of histone deacetylase inhibitors |
EP3508581A1 (en) | 2018-01-03 | 2019-07-10 | Kws Saat Se | Regeneration of genetically modified plants |
WO2019140278A1 (en) | 2018-01-11 | 2019-07-18 | Fred Hutchinson Cancer Research Center | Immunotherapy targeting core binding factor antigens |
IL257225A (en) | 2018-01-29 | 2018-04-09 | Yeda Res & Dev | Treatment of sarcoma |
EP3749350A4 (en) | 2018-02-08 | 2021-12-01 | Sangamo Therapeutics, Inc. | SPECIFICALLY MODIFIED TARGET-SPECIFIC NUCLEASES |
EP3752601A4 (en) | 2018-02-15 | 2022-03-23 | Memorial Sloan-Kettering Cancer Center | ANTI-FOXP3 DRUG COMPOSITIONS AND METHODS OF USE FOR ADOPTIVE CELL THERAPY |
WO2019161151A1 (en) | 2018-02-15 | 2019-08-22 | Monsanto Technology Llc | Compositions and methods for improving crop yields through trait stacking |
US20210079061A1 (en) | 2018-02-26 | 2021-03-18 | Fred Hutchinson Cancer Research Center | Compositions and methods for cellular immunotherapy |
RU2020133851A (ru) | 2018-03-16 | 2022-04-19 | Иммьюсофт Корпорейшн | B-клетки, генетически модифицированные для секреции фоллистатина, и способы их применения для лечения связанных с фоллистатином заболеваний, состояний, нарушений и для увеличения роста и силы мышц |
JP7334178B2 (ja) | 2018-03-19 | 2023-08-28 | リジェネロン・ファーマシューティカルズ・インコーポレイテッド | CRISPR/Cas系を使用した動物での転写モジュレーション |
EP3545756A1 (en) | 2018-03-28 | 2019-10-02 | KWS SAAT SE & Co. KGaA | Regeneration of plants in the presence of inhibitors of the histone methyltransferase ezh2 |
CN112585276A (zh) | 2018-04-05 | 2021-03-30 | 朱诺治疗学股份有限公司 | 产生表达重组受体的细胞的方法和相关组合物 |
CA3095027A1 (en) | 2018-04-05 | 2019-10-10 | Juno Therapeutics, Inc. | T cell receptors and engineered cells expressing same |
CA3095084A1 (en) | 2018-04-05 | 2019-10-10 | Juno Therapeutics, Inc. | T cells expressing a recombinant receptor, related polynucleotides and methods |
EP3784254A1 (en) | 2018-04-27 | 2021-03-03 | Iovance Biotherapeutics, Inc. | Closed process for expansion and gene editing of tumor infiltrating lymphocytes and uses of same in immunotherapy |
EP3567111A1 (en) | 2018-05-09 | 2019-11-13 | KWS SAAT SE & Co. KGaA | Gene for resistance to a pathogen of the genus heterodera |
US11690921B2 (en) | 2018-05-18 | 2023-07-04 | Sangamo Therapeutics, Inc. | Delivery of target specific nucleases |
GB201809273D0 (en) | 2018-06-06 | 2018-07-25 | Vib Vzw | Novel mutant plant cinnamoyl-coa reductase proteins |
WO2019236943A2 (en) | 2018-06-07 | 2019-12-12 | The Brigham And Women's Hospital, Inc. | Methods for generating hematopoietic stem cells |
WO2019238911A1 (en) | 2018-06-15 | 2019-12-19 | KWS SAAT SE & Co. KGaA | Methods for improving genome engineering and regeneration in plant ii |
WO2019238832A1 (en) | 2018-06-15 | 2019-12-19 | Nunhems B.V. | Seedless watermelon plants comprising modifications in an abc transporter gene |
CN112566924A (zh) | 2018-06-15 | 2021-03-26 | 科沃施种子欧洲股份两合公司 | 改善植物中基因组工程化和再生的方法 |
EP3807299A1 (en) | 2018-06-15 | 2021-04-21 | KWS SAAT SE & Co. KGaA | Methods for enhancing genome engineering efficiency |
US20220306699A1 (en) * | 2018-06-27 | 2022-09-29 | Altius Institute For Biomedical Sciences | Nucleic Acid Binding Domains and Methods of Use Thereof |
WO2020006131A2 (en) * | 2018-06-27 | 2020-01-02 | Altius Institute For Biomedical Sciences | Nucleases for genome editing |
WO2020006132A1 (en) * | 2018-06-27 | 2020-01-02 | Altius Institute For Biomedical Sciences | Gapped and tunable repeat units for use in genome editing and gene regulation compositions |
EP3818075A1 (en) | 2018-07-04 | 2021-05-12 | Ukko Inc. | Methods of de-epitoping wheat proteins and use of same for the treatment of celiac disease |
WO2020018964A1 (en) | 2018-07-20 | 2020-01-23 | Fred Hutchinson Cancer Research Center | Compositions and methods for controlled expression of antigen-specific receptors |
EP3841113A1 (en) | 2018-08-22 | 2021-06-30 | Fred Hutchinson Cancer Research Center | Immunotherapy targeting kras or her2 antigens |
JP2021536229A (ja) | 2018-08-23 | 2021-12-27 | サンガモ セラピューティクス, インコーポレイテッド | 操作された標的特異的な塩基エディター |
AU2019329984A1 (en) | 2018-08-28 | 2021-03-11 | Fred Hutchinson Cancer Center | Methods and compositions for adoptive T cell therapy incorporating induced notch signaling |
EP3623379A1 (en) | 2018-09-11 | 2020-03-18 | KWS SAAT SE & Co. KGaA | Beet necrotic yellow vein virus (bnyvv)-resistance modifying gene |
JP2022500052A (ja) | 2018-09-18 | 2022-01-04 | サンガモ セラピューティクス, インコーポレイテッド | プログラム細胞死1(pd1)特異的ヌクレアーゼ |
US20240165232A1 (en) | 2018-09-24 | 2024-05-23 | Fred Hutchinson Cancer Research Center | Chimeric receptor proteins and uses thereof |
IL262658A (en) | 2018-10-28 | 2020-04-30 | Memorial Sloan Kettering Cancer Center | Prevention of age related clonal hematopoiesis and diseases associated therewith |
BR112021008573A2 (pt) | 2018-11-05 | 2021-10-26 | Iovance Biotherapeutics, Inc. | Métodos para expandir linfócitos infiltrantes de tumor em uma população terapêutica de linfócitos infiltrantes de tumor, para tratar um sujeito com câncer e de expansão de células t, população terapêutica de linfócitos infiltrantes de tumor, composição de linfócitos infiltrantes de tumor, saco de infusão estéril, preparação criopreservada, e, uso da composição de linfócitos infiltrantes de tumor. |
SG11202104630PA (en) | 2018-11-05 | 2021-06-29 | Iovance Biotherapeutics Inc | Selection of improved tumor reactive t-cells |
MX2021004775A (es) | 2018-11-05 | 2021-06-08 | Iovance Biotherapeutics Inc | Expansion de linfocitos infiltrantes de tumores (tils) usando inhibidores de la via de proteina cinasa b (akt). |
SG11202104663PA (en) | 2018-11-05 | 2021-06-29 | Iovance Biotherapeutics Inc | Treatment of nsclc patients refractory for anti-pd-1 antibody |
BR112021008973A8 (pt) | 2018-11-09 | 2023-05-02 | Juno Therapeutics Inc | Receptores de células t específicos para mesotelina e uso dos mesmos em imunoterapia |
US20220023348A1 (en) | 2018-11-28 | 2022-01-27 | Forty Seven, Inc. | Genetically modified hspcs resistant to ablation regime |
EA202191463A1 (ru) | 2018-11-28 | 2021-10-13 | Борд Оф Риджентс, Дзе Юниверсити Оф Техас Систем | Мультиплексное редактирование генома иммунных клеток для повышения функциональности и устойчивости к подавляющей среде |
AU2019386830A1 (en) | 2018-11-29 | 2021-06-24 | Board Of Regents, The University Of Texas System | Methods for ex vivo expansion of natural killer cells and use thereof |
GB201820109D0 (en) | 2018-12-11 | 2019-01-23 | Vib Vzw | Plants with a lignin trait and udp-glycosyltransferase mutation |
US20220193131A1 (en) | 2018-12-19 | 2022-06-23 | Iovance Biotherapeutics, Inc. | Methods of Expanding Tumor Infiltrating Lymphocytes Using Engineered Cytokine Receptor Pairs and Uses Thereof |
WO2020146805A1 (en) | 2019-01-11 | 2020-07-16 | Acuitas Therapeutics, Inc. | Lipids for lipid nanoparticle delivery of active agents |
WO2020157573A1 (en) | 2019-01-29 | 2020-08-06 | The University Of Warwick | Methods for enhancing genome engineering efficiency |
WO2020163017A1 (en) | 2019-02-06 | 2020-08-13 | Klogenix Llc | Dna binding proteins and uses thereof |
AU2019428629A1 (en) | 2019-02-06 | 2021-01-28 | Sangamo Therapeutics, Inc. | Method for the treatment of mucopolysaccharidosis type I |
CA3130618A1 (en) | 2019-02-20 | 2020-08-27 | Fred Hutchinson Cancer Research Center | Binding proteins specific for ras neoantigens and uses thereof |
JP2022522473A (ja) | 2019-03-01 | 2022-04-19 | アイオバンス バイオセラピューティクス,インコーポレイテッド | 液性腫瘍からの腫瘍浸潤リンパ球の拡大培養及びその治療的使用 |
MX2021010611A (es) | 2019-03-05 | 2021-11-12 | The State Of Israel Ministry Of Agriculture & Rural Development Agricultural Res Organization Aro Vo | Aves con genoma editado. |
CN114174327A (zh) | 2019-03-08 | 2022-03-11 | 黑曜石疗法公司 | 用于可调整调控的cd40l组合物和方法 |
US20220160764A1 (en) | 2019-03-11 | 2022-05-26 | Fred Hutchinson Cancer Research Center | High avidity wt1 t cell receptors and uses thereof |
EP3708651A1 (en) | 2019-03-12 | 2020-09-16 | KWS SAAT SE & Co. KGaA | Improving plant regeneration |
JP7461368B2 (ja) | 2019-03-18 | 2024-04-03 | リジェネロン・ファーマシューティカルズ・インコーポレイテッド | タウの播種または凝集の遺伝的修飾因子を同定するためのcrispr/casスクリーニングプラットフォーム |
US11781131B2 (en) | 2019-03-18 | 2023-10-10 | Regeneron Pharmaceuticals, Inc. | CRISPR/Cas dropout screening platform to reveal genetic vulnerabilities associated with tau aggregation |
MX2021011426A (es) | 2019-03-19 | 2022-03-11 | Broad Inst Inc | Metodos y composiciones para editar secuencias de nucleótidos. |
MX2021012152A (es) | 2019-04-02 | 2021-11-03 | Sangamo Therapeutics Inc | Metodos para el tratamiento de beta-talasemia. |
JP2022527809A (ja) | 2019-04-03 | 2022-06-06 | リジェネロン・ファーマシューティカルズ・インコーポレイテッド | 抗体コード配列をセーフハーバー遺伝子座に挿入するための方法および組成物 |
CN113766921A (zh) | 2019-04-23 | 2021-12-07 | 桑格摩生物治疗股份有限公司 | 染色体9开放阅读框72基因表达的调节物及其用途 |
TW202106699A (zh) | 2019-04-26 | 2021-02-16 | 美商愛德維仁生物科技公司 | 用於玻璃體內遞送之變異體aav蛋白殼 |
TW202106700A (zh) | 2019-04-26 | 2021-02-16 | 美商聖加莫治療股份有限公司 | Aav 工程化 |
BR112021021200A2 (pt) | 2019-05-01 | 2021-12-21 | Juno Therapeutics Inc | Células expressando um receptor quimérico de um locus cd247 modificado, polinucleotídeos relacionados e métodos |
MX2021013219A (es) | 2019-05-01 | 2022-02-17 | Juno Therapeutics Inc | Celulas que expresan un receptor recombinante de un locus de tgfbr2 modificado, polinucleotidos relacionados y metodos. |
US20220249559A1 (en) | 2019-05-13 | 2022-08-11 | Iovance Biotherapeutics, Inc. | Methods and compositions for selecting tumor infiltrating lymphocytes and uses of the same in immunotherapy |
SI3840767T1 (sl) * | 2019-05-29 | 2024-04-30 | Hubro Therapeutics As | Peptidi |
WO2020244759A1 (en) | 2019-06-05 | 2020-12-10 | Klemm & Sohn Gmbh & Co. Kg | New plants having a white foliage phenotype |
AU2020290509A1 (en) | 2019-06-14 | 2021-11-11 | Regeneron Pharmaceuticals, Inc. | Models of tauopathy |
EP3757219A1 (en) | 2019-06-28 | 2020-12-30 | KWS SAAT SE & Co. KGaA | Enhanced plant regeneration and transformation by using grf1 booster gene |
KR20220042139A (ko) | 2019-07-04 | 2022-04-04 | 유코 아이엔씨. | 에피토프 제거된 알파 글리아딘 및 복강병 및 글루텐 민감성의 관리를 위한 이의 용도 |
IL268111A (en) | 2019-07-16 | 2021-01-31 | Fainzilber Michael | Methods of treating pain |
CA3144687A1 (en) | 2019-07-18 | 2021-01-21 | Steven A. Goldman | Cell-type selective immunoprotection of cells |
JP2022542102A (ja) | 2019-07-23 | 2022-09-29 | ムネモ・セラピューティクス | Suv39h1欠損免疫細胞 |
US10501404B1 (en) | 2019-07-30 | 2019-12-10 | Factor Bioscience Inc. | Cationic lipids and transfection methods |
US20220267732A1 (en) | 2019-08-01 | 2022-08-25 | Sana Biotechnology, Inc. | Dux4 expressing cells and uses thereof |
CA3147717A1 (en) | 2019-08-20 | 2021-02-25 | Fred Hutchinson Cancer Research Center | T-cell immunotherapy specific for wt-1 |
AU2020336302A1 (en) | 2019-08-23 | 2022-03-03 | Sana Biotechnology, Inc. | CD24 expressing cells and uses thereof |
CN114616002A (zh) | 2019-09-13 | 2022-06-10 | 瑞泽恩制药公司 | 使用由脂质纳米颗粒递送的crispr/cas系统在动物中进行的转录调控 |
JP2022550608A (ja) | 2019-10-02 | 2022-12-02 | サンガモ セラピューティクス, インコーポレイテッド | プリオン病の治療のためのジンクフィンガータンパク質転写因子 |
EP4041897A1 (en) * | 2019-10-11 | 2022-08-17 | Basf Plant Science Company GmbH | Increase in corn transformation efficacy by the n-terminus of a tale |
JP2022553389A (ja) | 2019-10-25 | 2022-12-22 | アイオバンス バイオセラピューティクス,インコーポレイテッド | 腫瘍浸潤リンパ球の遺伝子編集及び免疫療法におけるその使用 |
IL270306A (en) | 2019-10-30 | 2021-05-31 | Yeda Res & Dev | Prevention and treatment of pre-myeloid and myeloid malignancies |
AU2020376048A1 (en) | 2019-11-01 | 2022-06-02 | Sangamo Therapeutics, Inc. | Compositions and methods for genome engineering |
US20210130828A1 (en) | 2019-11-01 | 2021-05-06 | Sangamo Therapeutics, Inc. | Gin recombinase variants |
MX2022005572A (es) | 2019-11-08 | 2022-06-09 | Regeneron Pharma | Estrategias crispr y aav para la terapia de retinosquisis juvenil ligada al cromosoma x. |
CA3157872A1 (en) | 2019-11-12 | 2021-05-20 | Otto Torjek | Gene for resistance to a pathogen of the genus heterodera |
WO2021108363A1 (en) | 2019-11-25 | 2021-06-03 | Regeneron Pharmaceuticals, Inc. | Crispr/cas-mediated upregulation of humanized ttr allele |
JP2023506734A (ja) | 2019-12-11 | 2023-02-20 | アイオバンス バイオセラピューティクス,インコーポレイテッド | 腫瘍浸潤リンパ球(til)の産生のためのプロセス及びそれを使用する方法 |
IL271656A (en) | 2019-12-22 | 2021-06-30 | Yeda Res & Dev | System and methods for identifying cells that have undergone genome editing |
BR112022013355A2 (pt) | 2020-01-08 | 2022-09-20 | Obsidian Therapeutics Inc | Célula modificada, molécula de ácido nucleico, vetor, primeiro polinucleotídeo e segundo polinucleotídeo, método de produção de uma célula modificada, métodos para tratar ou prevenir uma doença em um sujeito em necessidade, para introduzir uma célula modificada, para modificar geneticamente uma ou mais células, e, sistema para a expressão sintonizável |
WO2021152587A1 (en) | 2020-01-30 | 2021-08-05 | Yeda Research And Development Co. Ltd. | Treating acute liver disease with tlr-mik inhibitors |
AU2021232598A1 (en) | 2020-03-04 | 2022-09-08 | Regeneron Pharmaceuticals, Inc. | Methods and compositions for sensitization of tumor cells to immune therapy |
EP4125348A1 (en) | 2020-03-23 | 2023-02-08 | Regeneron Pharmaceuticals, Inc. | Non-human animals comprising a humanized ttr locus comprising a v30m mutation and methods of use |
BR112022019060A2 (pt) | 2020-03-25 | 2022-11-29 | Sana Biotechnology Inc | Células neurais hipoimunogênicas para o tratamento de distúrbios e condições neurológicas |
JP2023523855A (ja) | 2020-05-04 | 2023-06-07 | アイオバンス バイオセラピューティクス,インコーポレイテッド | 腫瘍浸潤リンパ球の製造方法及び免疫療法におけるその使用 |
EP4146798A1 (en) | 2020-05-04 | 2023-03-15 | Saliogen Therapeutics, Inc. | Transposition-based therapies |
EP4146284A1 (en) | 2020-05-06 | 2023-03-15 | Cellectis S.A. | Methods to genetically modify cells for delivery of therapeutic proteins |
CN115803435A (zh) | 2020-05-06 | 2023-03-14 | 塞勒克提斯公司 | 用于在细胞基因组中靶向插入外源序列的方法 |
JP2023525304A (ja) | 2020-05-08 | 2023-06-15 | ザ ブロード インスティテュート,インコーポレーテッド | 標的二本鎖ヌクレオチド配列の両鎖同時編集のための方法および組成物 |
CN115835873A (zh) | 2020-05-13 | 2023-03-21 | 朱诺治疗学股份有限公司 | 用于产生表达重组受体的供体分批细胞的方法 |
CN111518220A (zh) * | 2020-05-14 | 2020-08-11 | 重庆英茂盛业生物科技有限公司 | 一种融合蛋白及其设计方法 |
US20230190871A1 (en) | 2020-05-20 | 2023-06-22 | Sana Biotechnology, Inc. | Methods and compositions for treatment of viral infections |
US20230183751A1 (en) | 2020-05-21 | 2023-06-15 | Oxford Genetics Limited | Hdr enhancers |
GB202007577D0 (en) | 2020-05-21 | 2020-07-08 | Oxford Genetics Ltd | Hdr enhancers |
EP4153739A1 (en) | 2020-05-21 | 2023-03-29 | Oxford Genetics Limited | Hdr enhancers |
GB202007578D0 (en) | 2020-05-21 | 2020-07-08 | Univ Oxford Innovation Ltd | Hdr enhancers |
WO2021247836A1 (en) | 2020-06-03 | 2021-12-09 | Board Of Regents, The University Of Texas System | Methods for targeting shp-2 to overcome resistance |
CN116234558A (zh) | 2020-06-26 | 2023-06-06 | 朱诺治疗学有限公司 | 条件性地表达重组受体的工程化t细胞、相关多核苷酸和方法 |
CA3189338A1 (en) | 2020-07-16 | 2022-01-20 | Acuitas Therapeutics, Inc. | Cationic lipids for use in lipid nanoparticles |
WO2022023576A1 (en) | 2020-07-30 | 2022-02-03 | Institut Curie | Immune cells defective for socs1 |
AU2021325941A1 (en) | 2020-08-13 | 2023-03-09 | Sana Biotechnology, Inc. | Methods of treating sensitized patients with hypoimmunogenic cells, and associated methods and compositions |
EP4204545A2 (en) | 2020-08-25 | 2023-07-05 | Kite Pharma, Inc. | T cells with improved functionality |
WO2022066973A1 (en) | 2020-09-24 | 2022-03-31 | Fred Hutchinson Cancer Research Center | Immunotherapy targeting pbk or oip5 antigens |
WO2022066965A2 (en) | 2020-09-24 | 2022-03-31 | Fred Hutchinson Cancer Research Center | Immunotherapy targeting sox2 antigens |
CA3196599A1 (en) | 2020-09-25 | 2022-03-31 | Sangamo Therapeutics, Inc. | Zinc finger fusion proteins for nucleobase editing |
US20230374447A1 (en) | 2020-10-05 | 2023-11-23 | Protalix Ltd. | Dicer-like knock-out plant cells |
WO2022076606A1 (en) | 2020-10-06 | 2022-04-14 | Iovance Biotherapeutics, Inc. | Treatment of nsclc patients with tumor infiltrating lymphocyte therapies |
CA3195019A1 (en) | 2020-10-06 | 2022-04-14 | Maria Fardis | Treatment of nsclc patients with tumor infiltrating lymphocyte therapies |
WO2022076353A1 (en) | 2020-10-06 | 2022-04-14 | Fred Hutchinson Cancer Research Center | Compositions and methods for treating mage-a1-expressing disease |
EP4228637A1 (en) | 2020-10-15 | 2023-08-23 | Yeda Research and Development Co. Ltd | Method of treating myeloid malignancies |
KR20230090367A (ko) | 2020-11-04 | 2023-06-21 | 주노 쎄러퓨티크스 인코퍼레이티드 | 변형된 불변 cd3 이뮤노글로불린 슈퍼패밀리 쇄 유전자좌로부터 키메라 수용체를 발현하는 세포 및 관련 폴리뉴클레오티드 및 방법 |
US20220162625A1 (en) * | 2020-11-11 | 2022-05-26 | Monsanto Technology Llc | Methods to improve site-directed integration frequency |
WO2022104109A1 (en) | 2020-11-13 | 2022-05-19 | Catamaran Bio, Inc. | Genetically modified natural killer cells and methods of use thereof |
CN117042600A (zh) | 2020-11-16 | 2023-11-10 | 猪改良英国有限公司 | 具有经编辑的anp32基因的抗甲型流感动物 |
WO2022115611A1 (en) | 2020-11-25 | 2022-06-02 | Catamaran Bio, Inc. | Cellular therapeutics engineered with signal modulators and methods of use thereof |
JP2023550979A (ja) | 2020-11-26 | 2023-12-06 | ウッコ インコーポレイテッド | 改変された高分子量グルテニンサブユニット及びその使用 |
US20240002839A1 (en) | 2020-12-02 | 2024-01-04 | Decibel Therapeutics, Inc. | Crispr sam biosensor cell lines and methods of use thereof |
EP4259651A2 (en) | 2020-12-14 | 2023-10-18 | Fred Hutchinson Cancer Center | Compositions and methods for cellular immunotherapy |
WO2022133149A1 (en) | 2020-12-17 | 2022-06-23 | Iovance Biotherapeutics, Inc. | Treatment of cancers with tumor infiltrating lymphocytes |
IL279559A (en) | 2020-12-17 | 2022-07-01 | Yeda Res & Dev | Controlling the ubiquitination process of mlkl for disease treatment |
WO2022133140A1 (en) | 2020-12-17 | 2022-06-23 | Iovance Biotherapeutics, Inc. | Treatment with tumor infiltrating lymphocyte therapies in combination with ctla-4 and pd-1 inhibitors |
JP2024500804A (ja) | 2020-12-18 | 2024-01-10 | イェダ リサーチ アンド ディベロップメント カンパニー リミテッド | Chd2ハプロ不全の処置における使用のための組成物及びその同定方法 |
EP4019639A1 (en) | 2020-12-22 | 2022-06-29 | KWS SAAT SE & Co. KGaA | Promoting regeneration and transformation in beta vulgaris |
EP4019638A1 (en) | 2020-12-22 | 2022-06-29 | KWS SAAT SE & Co. KGaA | Promoting regeneration and transformation in beta vulgaris |
KR20230137900A (ko) | 2020-12-31 | 2023-10-05 | 사나 바이오테크놀로지, 인크. | Car-t 활성을 조정하기 위한 방법 및 조성물 |
WO2022165260A1 (en) | 2021-01-29 | 2022-08-04 | Iovance Biotherapeutics, Inc. | Methods of making modified tumor infiltrating lymphocytes and their use in adoptive cell therapy |
WO2022198141A1 (en) | 2021-03-19 | 2022-09-22 | Iovance Biotherapeutics, Inc. | Methods for tumor infiltrating lymphocyte (til) expansion related to cd39/cd69 selection and gene knockout in tils |
JP2024511420A (ja) | 2021-03-22 | 2024-03-13 | ジュノー セラピューティクス インコーポレイテッド | ウイルスベクター粒子の効力を評価する方法 |
CN118019546A (zh) | 2021-03-23 | 2024-05-10 | 艾欧凡斯生物治疗公司 | 肿瘤浸润淋巴细胞的cish基因编辑及其在免疫疗法中的用途 |
US20240207318A1 (en) | 2021-04-19 | 2024-06-27 | Yongliang Zhang | Chimeric costimulatory receptors, chemokine receptors, and the use of same in cellular immunotherapies |
US20240191186A1 (en) | 2021-05-05 | 2024-06-13 | FUJIFILM Cellular Dynamics, Inc. | Methods and compositions for ipsc-derived microglia |
AU2022272489A1 (en) | 2021-05-10 | 2023-11-30 | Yissum Research Development Company Of The Hebrew University Of Jerusalem Ltd. | Pharmaceutical compositions for treating neurological conditions |
EP4337787A1 (en) | 2021-05-14 | 2024-03-20 | Akoya Biosciences, Inc. | Amplification of rna detection signals in biological samples |
JP2024519029A (ja) | 2021-05-17 | 2024-05-08 | アイオバンス バイオセラピューティクス,インコーポレイテッド | Pd-1遺伝子編集された腫瘍浸潤リンパ球及び免疫療法におけるその使用 |
US20220389436A1 (en) | 2021-05-26 | 2022-12-08 | FUJIFILM Cellular Dynamics, Inc. | Methods to prevent rapid silencing of genes in pluripotent stem cells |
KR20240013135A (ko) | 2021-05-27 | 2024-01-30 | 사나 바이오테크놀로지, 인크. | 조작된 hla-e 또는 hla-g를 포함하는 저면역원성 세포 |
AU2022294310A1 (en) | 2021-06-13 | 2024-05-09 | Yissum Research Development Company Of The Hebrew University Of Jerusalem Ltd. | Method for reprogramming human cells |
EP4370544A2 (en) | 2021-07-14 | 2024-05-22 | Sana Biotechnology, Inc. | Altered expression of y chromosome-linked antigens in hypoimmunogenic cells |
WO2023288281A2 (en) | 2021-07-15 | 2023-01-19 | Fred Hutchinson Cancer Center | Chimeric polypeptides |
EP4377446A1 (en) | 2021-07-28 | 2024-06-05 | Iovance Biotherapeutics, Inc. | Treatment of cancer patients with tumor infiltrating lymphocyte therapies in combination with kras inhibitors |
WO2023010135A1 (en) | 2021-07-30 | 2023-02-02 | Tune Therapeutics, Inc. | Compositions and methods for modulating expression of methyl-cpg binding protein 2 (mecp2) |
EP4377459A2 (en) | 2021-07-30 | 2024-06-05 | Tune Therapeutics, Inc. | Compositions and methods for modulating expression of frataxin (fxn) |
EP4130028A1 (en) | 2021-08-03 | 2023-02-08 | Rhazes Therapeutics Ltd | Engineered tcr complex and methods of using same |
EP4380966A2 (en) | 2021-08-03 | 2024-06-12 | Genicity Limited | Engineered tcr complex and methods of using same |
WO2023014825A1 (en) | 2021-08-03 | 2023-02-09 | Akoya Biosciences, Inc. | Rna detection by selective labeling and amplification |
WO2023014922A1 (en) | 2021-08-04 | 2023-02-09 | The Regents Of The University Of Colorado, A Body Corporate | Lat activating chimeric antigen receptor t cells and methods of use thereof |
AU2022325231A1 (en) | 2021-08-11 | 2024-02-08 | Sana Biotechnology, Inc. | Genetically modified cells for allogeneic cell therapy to reduce complement-mediated inflammatory reactions |
CA3227613A1 (en) | 2021-08-11 | 2023-02-16 | William Dowdle | Inducible systems for altering gene expression in hypoimmunogenic cells |
AU2022326565A1 (en) | 2021-08-11 | 2024-02-08 | Sana Biotechnology, Inc. | Genetically modified cells for allogeneic cell therapy |
KR20240073006A (ko) | 2021-08-11 | 2024-05-24 | 사나 바이오테크놀로지, 인크. | 동종이계 세포 요법을 위한 유전자 변형된 1차 세포 |
WO2023019225A2 (en) | 2021-08-11 | 2023-02-16 | Sana Biotechnology, Inc. | Genetically modified cells for allogeneic cell therapy to reduce instant blood mediated inflammatory reactions |
IL311333A (en) | 2021-09-09 | 2024-05-01 | Iovance Biotherapeutics Inc | Processes for the production of TIL products using a strong PD-1 TALEN |
CA3231278A1 (en) | 2021-09-10 | 2023-03-16 | Deepika Rajesh | Compositions of induced pluripotent stem cell-derived cells and methods of use thereof |
US20230242922A1 (en) | 2021-09-13 | 2023-08-03 | Life Technologies Corporation | Gene editing tools |
CA3234612A1 (en) | 2021-10-14 | 2023-04-20 | Weedout Ltd. | Methods of weed control |
WO2023069790A1 (en) | 2021-10-22 | 2023-04-27 | Sana Biotechnology, Inc. | Methods of engineering allogeneic t cells with a transgene in a tcr locus and associated compositions and methods |
WO2023076880A1 (en) | 2021-10-25 | 2023-05-04 | Board Of Regents, The University Of Texas System | Foxo1-targeted therapy for the treatment of cancer |
AU2022379633A1 (en) | 2021-10-27 | 2024-04-11 | Regeneron Pharmaceuticals, Inc. | Compositions and methods for expressing factor ix for hemophilia b therapy |
US20230187042A1 (en) | 2021-10-27 | 2023-06-15 | Iovance Biotherapeutics, Inc. | Systems and methods for coordinating manufacturing of cells for patient-specific immunotherapy |
WO2023077053A2 (en) | 2021-10-28 | 2023-05-04 | Regeneron Pharmaceuticals, Inc. | Crispr/cas-related methods and compositions for knocking out c5 |
WO2023077050A1 (en) | 2021-10-29 | 2023-05-04 | FUJIFILM Cellular Dynamics, Inc. | Dopaminergic neurons comprising mutations and methods of use thereof |
WO2023081900A1 (en) | 2021-11-08 | 2023-05-11 | Juno Therapeutics, Inc. | Engineered t cells expressing a recombinant t cell receptor (tcr) and related systems and methods |
AU2022408167A1 (en) | 2021-12-08 | 2024-06-06 | Regeneron Pharmaceuticals, Inc. | Mutant myocilin disease model and uses thereof |
WO2023105244A1 (en) | 2021-12-10 | 2023-06-15 | Pig Improvement Company Uk Limited | Editing tmprss2/4 for disease resistance in livestock |
GB202118058D0 (en) | 2021-12-14 | 2022-01-26 | Univ Warwick | Methods to increase yields in crops |
AU2022421236A1 (en) | 2021-12-22 | 2024-06-27 | Sangamo Therapeutics, Inc. | Novel zinc finger fusion proteins for nucleobase editing |
CA3241438A1 (en) | 2021-12-23 | 2023-06-29 | Sana Biotechnology, Inc. | Chimeric antigen receptor (car) t cells for treating autoimmune disease and associated methods |
WO2023126458A1 (en) | 2021-12-28 | 2023-07-06 | Mnemo Therapeutics | Immune cells with inactivated suv39h1 and modified tcr |
WO2023129940A1 (en) | 2021-12-30 | 2023-07-06 | Regel Therapeutics, Inc. | Compositions for modulating expression of sodium voltage-gated channel alpha subunit 1 and uses thereof |
WO2023131616A1 (en) | 2022-01-05 | 2023-07-13 | Vib Vzw | Means and methods to increase abiotic stress tolerance in plants |
WO2023131637A1 (en) | 2022-01-06 | 2023-07-13 | Vib Vzw | Improved silage grasses |
WO2023137472A2 (en) | 2022-01-14 | 2023-07-20 | Tune Therapeutics, Inc. | Compositions, systems, and methods for programming t cell phenotypes through targeted gene repression |
WO2023137471A1 (en) | 2022-01-14 | 2023-07-20 | Tune Therapeutics, Inc. | Compositions, systems, and methods for programming t cell phenotypes through targeted gene activation |
WO2023144199A1 (en) | 2022-01-26 | 2023-08-03 | Vib Vzw | Plants having reduced levels of bitter taste metabolites |
WO2023147488A1 (en) | 2022-01-28 | 2023-08-03 | Iovance Biotherapeutics, Inc. | Cytokine associated tumor infiltrating lymphocytes compositions and methods |
WO2023150620A1 (en) | 2022-02-02 | 2023-08-10 | Regeneron Pharmaceuticals, Inc. | Crispr-mediated transgene insertion in neonatal cells |
WO2023150798A1 (en) | 2022-02-07 | 2023-08-10 | Regeneron Pharmaceuticals, Inc. | Compositions and methods for defining optimal treatment timeframes in lysosomal disease |
WO2023154578A1 (en) | 2022-02-14 | 2023-08-17 | Sana Biotechnology, Inc. | Methods of treating patients exhibiting a prior failed therapy with hypoimmunogenic cells |
WO2023158836A1 (en) | 2022-02-17 | 2023-08-24 | Sana Biotechnology, Inc. | Engineered cd47 proteins and uses thereof |
TW202340457A (zh) | 2022-02-28 | 2023-10-16 | 美商凱特製藥公司 | 同種異體治療細胞 |
WO2023173123A1 (en) | 2022-03-11 | 2023-09-14 | Sana Biotechnology, Inc. | Genetically modified cells and compositions and uses thereof |
WO2023196877A1 (en) | 2022-04-06 | 2023-10-12 | Iovance Biotherapeutics, Inc. | Treatment of nsclc patients with tumor infiltrating lymphocyte therapies |
WO2023201207A1 (en) | 2022-04-11 | 2023-10-19 | Tenaya Therapeutics, Inc. | Adeno-associated virus with engineered capsid |
WO2023201369A1 (en) | 2022-04-15 | 2023-10-19 | Iovance Biotherapeutics, Inc. | Til expansion processes using specific cytokine combinations and/or akti treatment |
WO2023220040A1 (en) | 2022-05-09 | 2023-11-16 | Synteny Therapeutics, Inc. | Erythroparvovirus with a modified capsid for gene therapy |
WO2023220035A1 (en) | 2022-05-09 | 2023-11-16 | Synteny Therapeutics, Inc. | Erythroparvovirus compositions and methods for gene therapy |
WO2023220043A1 (en) | 2022-05-09 | 2023-11-16 | Synteny Therapeutics, Inc. | Erythroparvovirus with a modified genome for gene therapy |
WO2023220608A1 (en) | 2022-05-10 | 2023-11-16 | Iovance Biotherapeutics, Inc. | Treatment of cancer patients with tumor infiltrating lymphocyte therapies in combination with an il-15r agonist |
EP4279085A1 (en) | 2022-05-20 | 2023-11-22 | Mnemo Therapeutics | Compositions and methods for treating a refractory or relapsed cancer or a chronic infectious disease |
WO2023240282A1 (en) | 2022-06-10 | 2023-12-14 | Umoja Biopharma, Inc. | Engineered stem cells and uses thereof |
WO2023250511A2 (en) | 2022-06-24 | 2023-12-28 | Tune Therapeutics, Inc. | Compositions, systems, and methods for reducing low-density lipoprotein through targeted gene repression |
WO2024006911A1 (en) | 2022-06-29 | 2024-01-04 | FUJIFILM Holdings America Corporation | Ipsc-derived astrocytes and methods of use thereof |
GB2621813A (en) | 2022-06-30 | 2024-02-28 | Univ Newcastle | Preventing disease recurrence in Mitochondrial replacement therapy |
WO2024015881A2 (en) | 2022-07-12 | 2024-01-18 | Tune Therapeutics, Inc. | Compositions, systems, and methods for targeted transcriptional activation |
WO2024013514A2 (en) | 2022-07-15 | 2024-01-18 | Pig Improvement Company Uk Limited | Gene edited livestock animals having coronavirus resistance |
WO2024026474A1 (en) | 2022-07-29 | 2024-02-01 | Regeneron Pharmaceuticals, Inc. | Compositions and methods for transferrin receptor (tfr)-mediated delivery to the brain and muscle |
WO2024040254A2 (en) | 2022-08-19 | 2024-02-22 | Tune Therapeutics, Inc. | Compositions, systems, and methods for regulation of hepatitis b virus through targeted gene repression |
WO2024055017A1 (en) | 2022-09-09 | 2024-03-14 | Iovance Biotherapeutics, Inc. | Processes for generating til products using pd-1/tigit talen double knockdown |
WO2024055018A1 (en) | 2022-09-09 | 2024-03-14 | Iovance Biotherapeutics, Inc. | Processes for generating til products using pd-1/tigit talen double knockdown |
WO2024064642A2 (en) | 2022-09-19 | 2024-03-28 | Tune Therapeutics, Inc. | Compositions, systems, and methods for modulating t cell function |
WO2024062138A1 (en) | 2022-09-23 | 2024-03-28 | Mnemo Therapeutics | Immune cells comprising a modified suv39h1 gene |
WO2024073606A1 (en) | 2022-09-28 | 2024-04-04 | Regeneron Pharmaceuticals, Inc. | Antibody resistant modified receptors to enhance cell-based therapies |
WO2024098027A1 (en) | 2022-11-04 | 2024-05-10 | Iovance Biotherapeutics, Inc. | Methods for tumor infiltrating lymphocyte (til) expansion related to cd39/cd103 selection |
US20240182561A1 (en) | 2022-11-04 | 2024-06-06 | Regeneron Pharmaceuticals, Inc. | Calcium voltage-gated channel auxiliary subunit gamma 1 (cacng1) binding proteins and cacng1-mediated delivery to skeletal muscle |
WO2024098024A1 (en) | 2022-11-04 | 2024-05-10 | Iovance Biotherapeutics, Inc. | Expansion of tumor infiltrating lymphocytes from liquid tumors and therapeutic uses thereof |
WO2024100604A1 (en) | 2022-11-09 | 2024-05-16 | Juno Therapeutics Gmbh | Methods for manufacturing engineered immune cells |
WO2024107765A2 (en) | 2022-11-14 | 2024-05-23 | Regeneron Pharmaceuticals, Inc. | Compositions and methods for fibroblast growth factor receptor 3-mediated delivery to astrocytes |
WO2024107670A1 (en) | 2022-11-16 | 2024-05-23 | Regeneron Pharmaceuticals, Inc. | Chimeric proteins comprising membrane bound il-12 with protease cleavable linkers |
WO2024112711A2 (en) | 2022-11-21 | 2024-05-30 | Iovance Biotherapeutics, Inc. | Methods for assessing proliferation potency of gene-edited t cells |
Citations (2)
Publication number | Priority date | Publication date | Assignee | Title |
---|---|---|---|---|
WO2010079430A1 (en) * | 2009-01-12 | 2010-07-15 | Ulla Bonas | Modular dna-binding domains and methods of use |
WO2011072246A2 (en) * | 2009-12-10 | 2011-06-16 | Regents Of The University Of Minnesota | Tal effector-mediated dna modification |
Family Cites Families (137)
Publication number | Priority date | Publication date | Assignee | Title |
---|---|---|---|---|
US4179337A (en) | 1973-07-20 | 1979-12-18 | Davis Frank F | Non-immunogenic polypeptides |
US4217344A (en) | 1976-06-23 | 1980-08-12 | L'oreal | Compositions containing aqueous dispersions of lipid spheres |
US4235871A (en) | 1978-02-24 | 1980-11-25 | Papahadjopoulos Demetrios P | Method of encapsulating biologically active materials in lipid vesicles |
US4186183A (en) | 1978-03-29 | 1980-01-29 | The United States Of America As Represented By The Secretary Of The Army | Liposome carriers in chemotherapy of leishmaniasis |
US4261975A (en) | 1979-09-19 | 1981-04-14 | Merck & Co., Inc. | Viral liposome particle |
US4485054A (en) | 1982-10-04 | 1984-11-27 | Lipoderm Pharmaceuticals Limited | Method of encapsulating biologically active materials in multilamellar lipid vesicles (MLV) |
US4535060A (en) | 1983-01-05 | 1985-08-13 | Calgene, Inc. | Inhibition resistant 5-enolpyruvyl-3-phosphoshikimate synthetase, production and use |
US4501728A (en) | 1983-01-06 | 1985-02-26 | Technology Unlimited, Inc. | Masking of liposomes from RES recognition |
US4761373A (en) | 1984-03-06 | 1988-08-02 | Molecular Genetics, Inc. | Herbicide resistance in plants |
US5049386A (en) | 1985-01-07 | 1991-09-17 | Syntex (U.S.A.) Inc. | N-ω,(ω-1)-dialkyloxy)- and N-(ω,(ω-1)-dialkenyloxy)Alk-1-YL-N,N,N-tetrasubstituted ammonium lipids and uses therefor |
US4897355A (en) | 1985-01-07 | 1990-01-30 | Syntex (U.S.A.) Inc. | N[ω,(ω-1)-dialkyloxy]- and N-[ω,(ω-1)-dialkenyloxy]-alk-1-yl-N,N,N-tetrasubstituted ammonium lipids and uses therefor |
US4946787A (en) | 1985-01-07 | 1990-08-07 | Syntex (U.S.A.) Inc. | N-(ω,(ω-1)-dialkyloxy)- and N-(ω,(ω-1)-dialkenyloxy)-alk-1-yl-N,N,N-tetrasubstituted ammonium lipids and uses therefor |
US4797368A (en) | 1985-03-15 | 1989-01-10 | The United States Of America As Represented By The Department Of Health And Human Services | Adeno-associated virus as eukaryotic expression vector |
US4683195A (en) | 1986-01-30 | 1987-07-28 | Cetus Corporation | Process for amplifying, detecting, and/or-cloning nucleic acid sequences |
US4774085A (en) | 1985-07-09 | 1988-09-27 | 501 Board of Regents, Univ. of Texas | Pharmaceutical administration systems containing a mixture of immunomodulators |
US4940835A (en) | 1985-10-29 | 1990-07-10 | Monsanto Company | Glyphosate-resistant plants |
CA1293460C (en) | 1985-10-07 | 1991-12-24 | Brian Lee Sauer | Site-specific recombination of dna in yeast |
US4810648A (en) | 1986-01-08 | 1989-03-07 | Rhone Poulenc Agrochimie | Haloarylnitrile degrading gene, its use, and cells containing the gene |
ES2018274T5 (es) | 1986-03-11 | 1996-12-16 | Plant Genetic Systems Nv | Celulas vegetales resistentes a los inhibidores de glutamina sintetasa, preparadas por ingenieria genetica. |
US4975374A (en) | 1986-03-18 | 1990-12-04 | The General Hospital Corporation | Expression of wild type and mutant glutamine synthetase in foreign hosts |
US5276268A (en) | 1986-08-23 | 1994-01-04 | Hoechst Aktiengesellschaft | Phosphinothricin-resistance gene, and its use |
US5273894A (en) | 1986-08-23 | 1993-12-28 | Hoechst Aktiengesellschaft | Phosphinothricin-resistance gene, and its use |
US5013659A (en) | 1987-07-27 | 1991-05-07 | E. I. Du Pont De Nemours And Company | Nucleic acid fragment encoding herbicide resistant plant acetolactate synthase |
US5422251A (en) | 1986-11-26 | 1995-06-06 | Princeton University | Triple-stranded nucleic acids |
US4837028A (en) | 1986-12-24 | 1989-06-06 | Liposome Technology, Inc. | Liposomes with enhanced circulation time |
US5006333A (en) | 1987-08-03 | 1991-04-09 | Ddi Pharmaceuticals, Inc. | Conjugates of superoxide dismutase coupled to high molecular weight polyalkylene glycols |
US5162602A (en) | 1988-11-10 | 1992-11-10 | Regents Of The University Of Minnesota | Corn plants tolerant to sethoxydim and haloxyfop herbicides |
US5176996A (en) | 1988-12-20 | 1993-01-05 | Baylor College Of Medicine | Method for making synthetic oligonucleotides which bind specifically to target sites on duplex DNA molecules, by forming a colinear triplex, the synthetic oligonucleotides and methods of use |
US5501967A (en) | 1989-07-26 | 1996-03-26 | Mogen International, N.V./Rijksuniversiteit Te Leiden | Process for the site-directed integration of DNA into the genome of plants |
US5484956A (en) | 1990-01-22 | 1996-01-16 | Dekalb Genetics Corporation | Fertile transgenic Zea mays plant comprising heterologous DNA encoding Bacillus thuringiensis endotoxin |
US5264618A (en) | 1990-04-19 | 1993-11-23 | Vical, Inc. | Cationic lipids for intracellular delivery of biologically active molecules |
AU7979491A (en) | 1990-05-03 | 1991-11-27 | Vical, Inc. | Intracellular delivery of biologically active substances by means of self-assembling lipid complexes |
US5204253A (en) | 1990-05-29 | 1993-04-20 | E. I. Du Pont De Nemours And Company | Method and apparatus for introducing biological substances into living cells |
US5173414A (en) | 1990-10-30 | 1992-12-22 | Applied Immune Sciences, Inc. | Production of recombinant adeno-associated virus vectors |
US5767366A (en) | 1991-02-19 | 1998-06-16 | Louisiana State University Board Of Supervisors, A Governing Body Of Louisiana State University Agricultural And Mechanical College | Mutant acetolactate synthase gene from Ararbidopsis thaliana for conferring imidazolinone resistance to crop plants |
US5420032A (en) | 1991-12-23 | 1995-05-30 | Universitge Laval | Homing endonuclease which originates from chlamydomonas eugametos and recognizes and cleaves a 15, 17 or 19 degenerate double stranded nucleotide sequence |
US5356802A (en) | 1992-04-03 | 1994-10-18 | The Johns Hopkins University | Functional domains in flavobacterium okeanokoites (FokI) restriction endonuclease |
US5436150A (en) | 1992-04-03 | 1995-07-25 | The Johns Hopkins University | Functional domains in flavobacterium okeanokoities (foki) restriction endonuclease |
US5792640A (en) | 1992-04-03 | 1998-08-11 | The Johns Hopkins University | General method to clone hybrid restriction endonucleases using lig gene |
US5487994A (en) | 1992-04-03 | 1996-01-30 | The Johns Hopkins University | Insertion and deletion mutants of FokI restriction endonuclease |
US5792632A (en) | 1992-05-05 | 1998-08-11 | Institut Pasteur | Nucleotide sequence encoding the enzyme I-SceI and the uses thereof |
US5587308A (en) | 1992-06-02 | 1996-12-24 | The United States Of America As Represented By The Department Of Health & Human Services | Modified adeno-associated virus vector capable of expression from a novel promoter |
EP0604662B1 (en) | 1992-07-07 | 2008-06-18 | Japan Tobacco Inc. | Method of transforming monocotyledon |
CA2154581C (en) | 1993-02-12 | 2008-12-16 | Srinivasan Chandrasegaran | Functional domains in flavobacterium okeanokoites (foki) restriction endonuclease |
US6242568B1 (en) | 1994-01-18 | 2001-06-05 | The Scripps Research Institute | Zinc finger protein derivatives and methods therefor |
CA2681922C (en) | 1994-01-18 | 2012-05-15 | The Scripps Research Institute | Zinc finger protein derivatives and methods therefor |
US6140466A (en) | 1994-01-18 | 2000-10-31 | The Scripps Research Institute | Zinc finger protein derivatives and methods therefor |
WO1995025809A1 (en) | 1994-03-23 | 1995-09-28 | Ohio University | Compacted nucleic acids and their delivery to cells |
US5585245A (en) | 1994-04-22 | 1996-12-17 | California Institute Of Technology | Ubiquitin-based split protein sensor |
US6808904B2 (en) | 1994-06-16 | 2004-10-26 | Syngenta Participations Ag | Herbicide-tolerant protox genes produced by DNA shuffling |
GB9824544D0 (en) | 1998-11-09 | 1999-01-06 | Medical Res Council | Screening system |
ATE407205T1 (de) | 1994-08-20 | 2008-09-15 | Gendaq Ltd | Verbesserung in bezug auf bindungsproteine bei der erkennung von dna |
US7285416B2 (en) | 2000-01-24 | 2007-10-23 | Gendaq Limited | Regulated gene expression in plants |
US6326166B1 (en) | 1995-12-29 | 2001-12-04 | Massachusetts Institute Of Technology | Chimeric DNA-binding proteins |
US5789538A (en) | 1995-02-03 | 1998-08-04 | Massachusetts Institute Of Technology | Zinc finger proteins with high affinity new DNA binding specificities |
US5853973A (en) | 1995-04-20 | 1998-12-29 | American Cyanamid Company | Structure based designed herbicide resistant products |
US6084155A (en) | 1995-06-06 | 2000-07-04 | Novartis Ag | Herbicide-tolerant protoporphyrinogen oxidase ("protox") genes |
US5928638A (en) | 1996-06-17 | 1999-07-27 | Systemix, Inc. | Methods for gene transfer |
US5925523A (en) | 1996-08-23 | 1999-07-20 | President & Fellows Of Harvard College | Intraction trap assay, reagents and uses thereof |
JPH10117776A (ja) | 1996-10-22 | 1998-05-12 | Japan Tobacco Inc | インディカイネの形質転換方法 |
GB9703369D0 (en) | 1997-02-18 | 1997-04-09 | Lindqvist Bjorn H | Process |
GB2338237B (en) | 1997-02-18 | 2001-02-28 | Actinova Ltd | In vitro peptide or protein expression library |
US6342345B1 (en) | 1997-04-02 | 2002-01-29 | The Board Of Trustees Of The Leland Stanford Junior University | Detection of molecular interactions by reporter subunit complementation |
GB9710809D0 (en) | 1997-05-23 | 1997-07-23 | Medical Res Council | Nucleic acid binding proteins |
GB9710807D0 (en) | 1997-05-23 | 1997-07-23 | Medical Res Council | Nucleic acid binding proteins |
US6410248B1 (en) | 1998-01-30 | 2002-06-25 | Massachusetts Institute Of Technology | General strategy for selecting high-affinity zinc finger proteins for diverse DNA target sites |
WO1999045132A1 (en) | 1998-03-02 | 1999-09-10 | Massachusetts Institute Of Technology | Poly zinc finger proteins with improved linkers |
GB9819693D0 (en) | 1998-09-10 | 1998-11-04 | Zeneca Ltd | Glyphosate formulation |
US7013219B2 (en) | 1999-01-12 | 2006-03-14 | Sangamo Biosciences, Inc. | Regulation of endogenous gene expression in cells using zinc finger proteins |
US6453242B1 (en) | 1999-01-12 | 2002-09-17 | Sangamo Biosciences, Inc. | Selection of sites for targeting by zinc finger proteins and methods of designing zinc finger proteins to bind to preselected sites |
US6599692B1 (en) | 1999-09-14 | 2003-07-29 | Sangamo Bioscience, Inc. | Functional genomics using zinc finger proteins |
US7070934B2 (en) | 1999-01-12 | 2006-07-04 | Sangamo Biosciences, Inc. | Ligand-controlled regulation of endogenous gene expression |
US6534261B1 (en) | 1999-01-12 | 2003-03-18 | Sangamo Biosciences, Inc. | Regulation of endogenous gene expression in cells using zinc finger proteins |
AU2982900A (en) | 1999-02-03 | 2000-08-25 | Children's Medical Center Corporation | Gene repair involving the induction of double-stranded dna cleavage at a chromosomal target site |
US6451732B1 (en) | 1999-06-04 | 2002-09-17 | Syngenta, Limited | Herbicidal compositions of glyphosate trimesium |
AU776576B2 (en) | 1999-12-06 | 2004-09-16 | Sangamo Biosciences, Inc. | Methods of using randomized libraries of zinc finger proteins for the identification of gene function |
EP1180480B1 (en) | 2000-01-26 | 2012-11-14 | Dai Nippon Printing Co., Ltd. | Heat-sealing method |
DE60143192D1 (de) | 2000-02-08 | 2010-11-18 | Sangamo Biosciences Inc | Zellen zur entdeckung von medikamenten |
US20020061512A1 (en) | 2000-02-18 | 2002-05-23 | Kim Jin-Soo | Zinc finger domains and methods of identifying same |
DE60126483T2 (de) | 2000-04-28 | 2007-12-06 | Sangamo BioSciences, Inc., Richmond | Gezielte Modifikation der Chromatinstruktur |
ATE374249T1 (de) | 2000-04-28 | 2007-10-15 | Sangamo Biosciences Inc | Datenbanken von regulatorischen sequenzen, methoden zu ihrer herstellung und verwendung |
WO2001088197A2 (en) | 2000-05-16 | 2001-11-22 | Massachusetts Institute Of Technology | Methods and compositions for interaction trap assays |
US6919204B2 (en) | 2000-09-29 | 2005-07-19 | Sangamo Biosciences, Inc. | Modulation of gene expression using localization domains |
US6368227B1 (en) | 2000-11-17 | 2002-04-09 | Steven Olson | Method of swinging on a swing |
WO2002044376A2 (en) | 2000-11-28 | 2002-06-06 | Sangamo Biosciences, Inc. | Modulation of gene expression using insulator binding proteins |
US7067317B2 (en) | 2000-12-07 | 2006-06-27 | Sangamo Biosciences, Inc. | Regulation of angiogenesis with zinc finger proteins |
US7273923B2 (en) | 2001-01-22 | 2007-09-25 | Sangamo Biosciences, Inc. | Zinc finger proteins for DNA binding and gene regulation in plants |
GB0108491D0 (en) | 2001-04-04 | 2001-05-23 | Gendaq Ltd | Engineering zinc fingers |
US7262054B2 (en) | 2002-01-22 | 2007-08-28 | Sangamo Biosciences, Inc. | Zinc finger proteins for DNA binding and gene regulation in plants |
ATE347593T1 (de) | 2002-01-23 | 2006-12-15 | Univ Utah Res Found | Zielgerichtete chromosomale mutagenese mit zinkfingernukleasen |
WO2009095742A1 (en) | 2008-01-31 | 2009-08-06 | Cellectis | New i-crei derived single-chain meganuclease and uses thereof |
US20060078552A1 (en) | 2002-03-15 | 2006-04-13 | Sylvain Arnould | Hybrid and single chain meganucleases and use thereof |
WO2003080809A2 (en) | 2002-03-21 | 2003-10-02 | Sangamo Biosciences, Inc. | Methods and compositions for using zinc finger endonucleases to enhance homologous recombination |
US7361635B2 (en) | 2002-08-29 | 2008-04-22 | Sangamo Biosciences, Inc. | Simultaneous modulation of multiple genes |
EP1581610A4 (en) | 2002-09-05 | 2009-05-27 | California Inst Of Techn | USE OF CHIMERIC NUCLEASES TO STIMULATE GENE TARGETING |
JP4966006B2 (ja) | 2003-01-28 | 2012-07-04 | セレクティス | カスタムメイドメガヌクレアーゼおよびその使用 |
US7888121B2 (en) | 2003-08-08 | 2011-02-15 | Sangamo Biosciences, Inc. | Methods and compositions for targeted cleavage and recombination |
US8409861B2 (en) | 2003-08-08 | 2013-04-02 | Sangamo Biosciences, Inc. | Targeted deletion of cellular DNA sequences |
US7972854B2 (en) | 2004-02-05 | 2011-07-05 | Sangamo Biosciences, Inc. | Methods and compositions for targeted cleavage and recombination |
US7189691B2 (en) | 2004-04-01 | 2007-03-13 | The Administrators Of The Tulane Educational Fund | Methods and compositions for treating leukemia |
US20060063231A1 (en) | 2004-09-16 | 2006-03-23 | Sangamo Biosciences, Inc. | Compositions and methods for protein production |
CN101273141B (zh) | 2005-07-26 | 2013-03-27 | 桑格摩生物科学股份有限公司 | 外源核酸序列的靶向整合和表达 |
EP2341135A3 (en) | 2005-10-18 | 2011-10-12 | Precision Biosciences | Rationally-designed meganucleases with altered sequence specificity and DNA-binding affinity |
WO2007060495A1 (en) | 2005-10-25 | 2007-05-31 | Cellectis | I-crei homing endonuclease variants having novel cleavage specificity and use thereof |
AU2007254251A1 (en) | 2006-05-19 | 2007-11-29 | Sangamo Therapeutics, Inc. | Methods and compositions for inactivation of dihydrofolate reductase |
WO2007139898A2 (en) | 2006-05-25 | 2007-12-06 | Sangamo Biosciences, Inc. | Variant foki cleavage half-domains |
JP5551432B2 (ja) | 2006-05-25 | 2014-07-16 | サンガモ バイオサイエンシーズ, インコーポレイテッド | 遺伝子不活性化のための方法と組成物 |
WO2008010009A1 (en) | 2006-07-18 | 2008-01-24 | Cellectis | Meganuclease variants cleaving a dna target sequence from a rag gene and uses thereof |
EP3070169B1 (en) | 2006-12-14 | 2018-05-09 | Dow AgroSciences LLC | Optimized non-canonical zinc finger proteins |
ATE489465T1 (de) | 2007-04-26 | 2010-12-15 | Sangamo Biosciences Inc | Gezielte integration in die ppp1r12c-position |
US8790345B2 (en) | 2007-08-21 | 2014-07-29 | Zimmer, Inc. | Titanium alloy with oxidized zirconium for a prosthetic implant |
CN101878307B (zh) | 2007-09-27 | 2017-07-28 | 陶氏益农公司 | 以5‑烯醇式丙酮酰莽草酸‑3‑磷酸合酶基因为靶的改造锌指蛋白 |
US8563314B2 (en) | 2007-09-27 | 2013-10-22 | Sangamo Biosciences, Inc. | Methods and compositions for modulating PD1 |
WO2009042753A1 (en) | 2007-09-28 | 2009-04-02 | Two Blades Foundation | Bs3 resistance gene and methods of use |
RU2499599C2 (ru) | 2007-09-28 | 2013-11-27 | Интрексон Корпорейшн | Конструкции терапевтического переключения генов и биореакторы для экспрессии биотерапевтических молекул и их применение |
CA2703045C (en) | 2007-10-25 | 2017-02-14 | Sangamo Biosciences, Inc. | Methods and compositions for targeted integration |
AU2009313275B2 (en) | 2008-11-10 | 2015-08-20 | Two Blades Foundation | Pathogen-inducible promoters and their use in enhancing the disease resistance of plants |
US9834787B2 (en) | 2009-04-09 | 2017-12-05 | Sangamo Therapeutics, Inc. | Targeted integration into stem cells |
US8772008B2 (en) | 2009-05-18 | 2014-07-08 | Sangamo Biosciences, Inc. | Methods and compositions for increasing nuclease activity |
JP2012529287A (ja) | 2009-06-11 | 2012-11-22 | トゥールゲン インコーポレイション | 部位−特異的ヌクレアーゼを用いた標的ゲノムの再配列 |
EP2449135B1 (en) | 2009-06-30 | 2016-01-06 | Sangamo BioSciences, Inc. | Rapid screening of biologically active nucleases and isolation of nuclease-modified cells |
WO2011017293A2 (en) | 2009-08-03 | 2011-02-10 | The General Hospital Corporation | Engineering of zinc finger arrays by context-dependent assembly |
US8586526B2 (en) * | 2010-05-17 | 2013-11-19 | Sangamo Biosciences, Inc. | DNA-binding proteins and uses thereof |
AU2010282958B2 (en) | 2009-08-11 | 2015-05-21 | Sangamo Therapeutics, Inc. | Organisms homozygous for targeted modification |
WO2011049627A1 (en) | 2009-10-22 | 2011-04-28 | Dow Agrosciences Llc | Engineered zinc finger proteins targeting plant genes involved in fatty acid biosynthesis |
US8956828B2 (en) | 2009-11-10 | 2015-02-17 | Sangamo Biosciences, Inc. | Targeted disruption of T cell receptor genes using engineered zinc finger protein nucleases |
CN105384827A (zh) | 2009-11-27 | 2016-03-09 | 巴斯夫植物科学有限公司 | 嵌合内切核酸酶及其用途 |
DE112010004584T5 (de) | 2009-11-27 | 2012-11-29 | Basf Plant Science Company Gmbh | Chimäre Endonukleasen und Anwendungen davon |
CA2782014C (en) | 2009-11-27 | 2021-08-31 | Basf Plant Science Company Gmbh | Optimized endonucleases and uses thereof |
US20110203012A1 (en) | 2010-01-21 | 2011-08-18 | Dotson Stanton B | Methods and compositions for use of directed recombination in plant breeding |
EP4328304A2 (en) | 2010-02-08 | 2024-02-28 | Sangamo Therapeutics, Inc. | Engineered cleavage half-domains |
EP2660318A1 (en) | 2010-02-09 | 2013-11-06 | Sangamo BioSciences, Inc. | Targeted genomic modification with partially single-stranded donor molecules |
US9567573B2 (en) | 2010-04-26 | 2017-02-14 | Sangamo Biosciences, Inc. | Genome editing of a Rosa locus using nucleases |
EP2392208B1 (en) | 2010-06-07 | 2016-05-04 | Helmholtz Zentrum München Deutsches Forschungszentrum für Gesundheit und Umwelt (GmbH) | Fusion proteins comprising a DNA-binding domain of a Tal effector protein and a non-specific cleavage domain of a restriction nuclease and their use |
CA2802360A1 (en) | 2010-06-14 | 2011-12-22 | Iowa State University Research Foundation, Inc. | Nuclease activity of tal effector and foki fusion protein |
US9405700B2 (en) | 2010-11-04 | 2016-08-02 | Sonics, Inc. | Methods and apparatus for virtualization in an integrated circuit |
WO2013074999A1 (en) * | 2011-11-16 | 2013-05-23 | Sangamo Biosciences, Inc. | Modified dna-binding proteins and uses thereof |
-
2011
- 2011-05-17 US US13/068,735 patent/US8586526B2/en active Active
- 2011-05-17 JP JP2013511148A patent/JP6208580B2/ja active Active
- 2011-05-17 CA CA2798988A patent/CA2798988C/en active Active
- 2011-05-17 EP EP11783865.6A patent/EP2571512B1/en active Active
- 2011-05-17 EP EP16200664.7A patent/EP3156062A1/en not_active Ceased
- 2011-05-17 CN CN201180034243.1A patent/CN103025344B/zh active Active
- 2011-05-17 WO PCT/US2011/000885 patent/WO2011146121A1/en active Application Filing
- 2011-05-17 KR KR1020127032393A patent/KR101953237B1/ko active IP Right Grant
- 2011-05-17 AU AU2011256838A patent/AU2011256838B2/en active Active
-
2012
- 2012-11-11 IL IL222961A patent/IL222961B/en active IP Right Grant
-
2013
- 2013-10-28 US US14/065,028 patent/US9493750B2/en active Active
- 2013-10-28 US US14/064,991 patent/US9322005B2/en active Active
- 2013-10-28 US US14/065,055 patent/US8912138B2/en active Active
-
2016
- 2016-07-08 JP JP2016135751A patent/JP2016182143A/ja active Pending
- 2016-10-03 US US15/284,164 patent/US9783827B2/en active Active
-
2017
- 2017-09-20 US US15/709,969 patent/US10253333B2/en active Active
-
2019
- 2019-02-12 US US16/274,024 patent/US11661612B2/en active Active
-
2021
- 2021-06-11 US US17/346,020 patent/US20220356493A1/en active Pending
Patent Citations (2)
Publication number | Priority date | Publication date | Assignee | Title |
---|---|---|---|---|
WO2010079430A1 (en) * | 2009-01-12 | 2010-07-15 | Ulla Bonas | Modular dna-binding domains and methods of use |
WO2011072246A2 (en) * | 2009-12-10 | 2011-06-16 | Regents Of The University Of Minnesota | Tal effector-mediated dna modification |
Non-Patent Citations (9)
Title |
---|
BOCH,J. ET AL.: "Breaking the code of DNA binding specificity of TAL-type III effectors.", SCIENCE, vol. 326, no. 5959, JPN6014046456, 29 October 2009 (2009-10-29), pages 1509 - 1512, XP002581597, ISSN: 0003105320 * |
KACHALOVA, G.S. ET AL: "Structural Analysis of the Heterodimeric Type IIS Restriction Endonuclease R.BspD6I Acting as a Comp", J. MOL. BIOL., vol. 384, no. 2, JPN6015025974, 2008, pages 489 - 502, XP025609976, ISSN: 0003105328, DOI: 10.1016/j.jmb.2008.09.033 * |
LIPPOW, S.M. ET AL.: "Creation of a type IIS restriction endonuclease with a long recognition sequence.", NUCLEIC ACIDS RES., vol. 37, no. 9, JPN7015001765, 2009, pages 3061 - 3073, XP008155323, ISSN: 0003105323, DOI: 10.1093/nar/gkp182 * |
MOSCOU, M.J. ET AL.: "A simple cipher governs DNA recognition by TAL effectors.", SCIENCE, vol. 326, no. 5959, JPN6015025970, 2009, pages 1501, XP002599998, ISSN: 0003105325 * |
SHUKLA,V.K. ET AL.: "Precise genome modification in the crop species Zea mays using zinc-finger nucleases.", NATURE, vol. 459, no. 7245, JPN6014046474, 21 May 2009 (2009-05-21), pages 437 - 441, ISSN: 0003105324 * |
SZUREK, B. ET AL.: "Eukaryotic features of the Xanthomonas type III effector AvrBs3: protein domains involved in transcr", PLANT J., vol. 26, no. 5, JPN6015025968, 2001, pages 523 - 534, ISSN: 0003105322 * |
TOWNSEND,J.A. ET AL.: "High-frequency modification of plant genes using engineered zinc-finger nucleases.", NATURE, vol. 459, no. 7245, JPN6014046476, 21 May 2009 (2009-05-21), pages 442 - 445, XP002588972, ISSN: 0003105326, DOI: 10.1038/NATURE07845 * |
WILLIAMS, R.J.: "Restriction endonucleases: classification, properties, and applications.", MOL BIOTECHNOL., vol. 23, no. 3, JPN6015025973, 2003, pages 225 - 243, XP009034361, ISSN: 0003105327, DOI: 10.1385/MB:23:3:225 * |
ZHU, W. ET AL.: "The C terminus of AvrXa10 can be replaced by the transcriptional activation domain of VP16 from the", PLANT CELL, vol. 11, no. 9, JPN6015025967, 1999, pages 1665 - 1674, ISSN: 0003105321 * |
Cited By (33)
Publication number | Priority date | Publication date | Assignee | Title |
---|---|---|---|---|
US10920242B2 (en) | 2011-02-25 | 2021-02-16 | Recombinetics, Inc. | Non-meiotic allele introgression |
US11472849B2 (en) | 2011-07-15 | 2022-10-18 | The General Hospital Corporation | Methods of transcription activator like effector assembly |
JP2014521317A (ja) * | 2011-07-15 | 2014-08-28 | ザ ジェネラル ホスピタル コーポレイション | 転写活性化因子様エフェクターの組立て方法 |
US10273271B2 (en) | 2011-07-15 | 2019-04-30 | The General Hospital Corporation | Methods of transcription activator like effector assembly |
US9890364B2 (en) | 2012-05-29 | 2018-02-13 | The General Hospital Corporation | TAL-Tet1 fusion proteins and methods of use thereof |
US10894950B2 (en) | 2012-05-29 | 2021-01-19 | The General Hospital Corporation | TAL-Tet1 fusion proteins and methods of use thereof |
US11891631B2 (en) | 2012-10-12 | 2024-02-06 | The General Hospital Corporation | Transcription activator-like effector (tale) - lysine-specific demethylase 1 (LSD1) fusion proteins |
US10731167B2 (en) | 2013-02-07 | 2020-08-04 | The General Hospital Corporation | Tale transcriptional activators |
US10676749B2 (en) | 2013-02-07 | 2020-06-09 | The General Hospital Corporation | Tale transcriptional activators |
JP6993063B2 (ja) | 2013-07-26 | 2022-01-13 | プレジデント アンド フェローズ オブ ハーバード カレッジ | ゲノムエンジニアリング |
JP2020110152A (ja) * | 2013-07-26 | 2020-07-27 | プレジデント アンド フェローズ オブ ハーバード カレッジ | ゲノムエンジニアリング |
US11306328B2 (en) | 2013-07-26 | 2022-04-19 | President And Fellows Of Harvard College | Genome engineering |
JP2015033365A (ja) * | 2013-08-09 | 2015-02-19 | 国立大学法人広島大学 | Dna結合ドメインを含むポリペプチド |
US10006011B2 (en) | 2013-08-09 | 2018-06-26 | Hiroshima University | Polypeptide containing DNA-binding domain |
US10030235B2 (en) | 2013-08-09 | 2018-07-24 | Hiroshima University | Polypeptide containing DNA-binding domain |
WO2015020218A1 (ja) * | 2013-08-09 | 2015-02-12 | 独立行政法人理化学研究所 | Dna結合タンパク質ドメインの改変による高活性taleタンパク質 |
WO2015019672A1 (ja) * | 2013-08-09 | 2015-02-12 | 国立大学法人広島大学 | Dna結合ドメインを含むポリペプチド |
US11359186B2 (en) | 2013-08-09 | 2022-06-14 | Hiroshima University | Polypeptide containing DNA-binding domain |
JP2020089386A (ja) * | 2013-08-27 | 2020-06-11 | リコンビネティクス・インコーポレイテッドRecombinetics,Inc. | 効率的な非減数分裂対立遺伝子移入 |
US11477969B2 (en) | 2013-08-27 | 2022-10-25 | Recombinetics, Inc. | Efficient non-meiotic allele introgression in livestock |
JP2016529904A (ja) * | 2013-08-27 | 2016-09-29 | リコンビネティクス・インコーポレイテッドRecombinetics,Inc. | 効率的な非減数分裂対立遺伝子移入 |
US10959414B2 (en) | 2013-08-27 | 2021-03-30 | Recombinetics, Inc. | Efficient non-meiotic allele introgression |
WO2015068785A1 (ja) | 2013-11-06 | 2015-05-14 | 国立大学法人広島大学 | 核酸挿入用ベクター |
EP3865575A1 (en) | 2013-11-06 | 2021-08-18 | Hiroshima University | Vector for nucleic acid insertion |
JP2018057407A (ja) * | 2013-11-18 | 2018-04-12 | クリスパー セラピューティクス アーゲー | Crispr−casシステムの材料及び方法 |
WO2016159110A1 (ja) * | 2015-03-31 | 2016-10-06 | 国立研究開発法人農業生物資源研究所 | バイナリー遺伝子発現システム |
JPWO2016159110A1 (ja) * | 2015-03-31 | 2018-02-01 | 国立研究開発法人農業・食品産業技術総合研究機構 | バイナリー遺伝子発現システム |
JP2016175922A (ja) * | 2016-04-25 | 2016-10-06 | 国立大学法人広島大学 | Dna結合ドメインを含むポリペプチド |
JP2020531572A (ja) * | 2017-08-04 | 2020-11-05 | 北京大学Peking University | メチル化により修飾されたdna塩基を特異的に認識するtale rvdおよびその使用 |
JP7207665B2 (ja) | 2017-08-04 | 2023-01-18 | 北京大学 | メチル化により修飾されたdna塩基を特異的に認識するtale rvdおよびその使用 |
US11897920B2 (en) | 2017-08-04 | 2024-02-13 | Peking University | Tale RVD specifically recognizing DNA base modified by methylation and application thereof |
US11624077B2 (en) | 2017-08-08 | 2023-04-11 | Peking University | Gene knockout method |
WO2020209332A1 (ja) | 2019-04-09 | 2020-10-15 | 国立研究開発法人科学技術振興機構 | 核酸結合性タンパク質 |
Also Published As
Publication number | Publication date |
---|---|
US20190169640A1 (en) | 2019-06-06 |
US8586526B2 (en) | 2013-11-19 |
US9322005B2 (en) | 2016-04-26 |
US20170016030A1 (en) | 2017-01-19 |
US20140134741A1 (en) | 2014-05-15 |
IL222961B (en) | 2018-08-30 |
US9783827B2 (en) | 2017-10-10 |
AU2011256838A1 (en) | 2012-12-06 |
IL222961A0 (en) | 2013-02-03 |
WO2011146121A1 (en) | 2011-11-24 |
CN103025344A (zh) | 2013-04-03 |
AU2011256838B2 (en) | 2014-10-09 |
CN103025344B (zh) | 2016-06-29 |
EP2571512B1 (en) | 2017-08-23 |
US11661612B2 (en) | 2023-05-30 |
JP6208580B2 (ja) | 2017-10-04 |
US20220356493A1 (en) | 2022-11-10 |
US10253333B2 (en) | 2019-04-09 |
CA2798988C (en) | 2020-03-10 |
CA2798988A1 (en) | 2011-11-24 |
KR101953237B1 (ko) | 2019-02-28 |
US20140134740A1 (en) | 2014-05-15 |
JP2016182143A (ja) | 2016-10-20 |
US20110301073A1 (en) | 2011-12-08 |
EP3156062A1 (en) | 2017-04-19 |
KR20130111219A (ko) | 2013-10-10 |
EP2571512A4 (en) | 2013-11-20 |
US20140134723A1 (en) | 2014-05-15 |
US8912138B2 (en) | 2014-12-16 |
US20180010152A1 (en) | 2018-01-11 |
US9493750B2 (en) | 2016-11-15 |
EP2571512A1 (en) | 2013-03-27 |
Similar Documents
Publication | Publication Date | Title |
---|---|---|
US11661612B2 (en) | DNA-binding proteins and uses thereof | |
JP6812478B2 (ja) | ヌクレアーゼ媒介性遺伝子破壊を増強するための方法および組成物 | |
JP6144691B2 (ja) | 修飾されたdna結合タンパク質およびその使用 | |
JP5940977B2 (ja) | 標的改変によるホモ接合生物 | |
CA2910427A1 (en) | Delivery methods and compositions for nuclease-mediated genome engineering |
Legal Events
Date | Code | Title | Description |
---|---|---|---|
A621 | Written request for application examination |
Free format text: JAPANESE INTERMEDIATE CODE: A621 Effective date: 20140509 |
|
A131 | Notification of reasons for refusal |
Free format text: JAPANESE INTERMEDIATE CODE: A131 Effective date: 20150630 |
|
A521 | Request for written amendment filed |
Free format text: JAPANESE INTERMEDIATE CODE: A523 Effective date: 20150930 |
|
A02 | Decision of refusal |
Free format text: JAPANESE INTERMEDIATE CODE: A02 Effective date: 20160308 |
|
A521 | Request for written amendment filed |
Free format text: JAPANESE INTERMEDIATE CODE: A523 Effective date: 20160708 |
|
A521 | Request for written amendment filed |
Free format text: JAPANESE INTERMEDIATE CODE: A821 Effective date: 20160712 |
|
A911 | Transfer to examiner for re-examination before appeal (zenchi) |
Free format text: JAPANESE INTERMEDIATE CODE: A911 Effective date: 20160804 |
|
A912 | Re-examination (zenchi) completed and case transferred to appeal board |
Free format text: JAPANESE INTERMEDIATE CODE: A912 Effective date: 20161007 |
|
A521 | Request for written amendment filed |
Free format text: JAPANESE INTERMEDIATE CODE: A523 Effective date: 20170531 |
|
A61 | First payment of annual fees (during grant procedure) |
Free format text: JAPANESE INTERMEDIATE CODE: A61 Effective date: 20170907 |
|
R150 | Certificate of patent or registration of utility model |
Ref document number: 6208580 Country of ref document: JP Free format text: JAPANESE INTERMEDIATE CODE: R150 |
|
R250 | Receipt of annual fees |
Free format text: JAPANESE INTERMEDIATE CODE: R250 |
|
R250 | Receipt of annual fees |
Free format text: JAPANESE INTERMEDIATE CODE: R250 |
|
R250 | Receipt of annual fees |
Free format text: JAPANESE INTERMEDIATE CODE: R250 |
|
R250 | Receipt of annual fees |
Free format text: JAPANESE INTERMEDIATE CODE: R250 |