EP1240513A2 - High throughput screening (hts) assays for protein variants with reduced antibody binding capacity - Google Patents
High throughput screening (hts) assays for protein variants with reduced antibody binding capacityInfo
- Publication number
- EP1240513A2 EP1240513A2 EP00981185A EP00981185A EP1240513A2 EP 1240513 A2 EP1240513 A2 EP 1240513A2 EP 00981185 A EP00981185 A EP 00981185A EP 00981185 A EP00981185 A EP 00981185A EP 1240513 A2 EP1240513 A2 EP 1240513A2
- Authority
- EP
- European Patent Office
- Prior art keywords
- protein
- antibodies
- library
- antibody
- variants
- Prior art date
- Legal status (The legal status is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the status listed.)
- Ceased
Links
- 108090000623 proteins and genes Proteins 0.000 title claims abstract description 261
- 102000004169 proteins and genes Human genes 0.000 title claims abstract description 219
- 230000027455 binding Effects 0.000 title claims abstract description 84
- 238000003556 assay Methods 0.000 title claims abstract description 33
- 230000002829 reductive effect Effects 0.000 title claims abstract description 27
- 238000013537 high throughput screening Methods 0.000 title description 14
- 238000000034 method Methods 0.000 claims abstract description 177
- 238000012216 screening Methods 0.000 claims abstract description 20
- 238000012258 culturing Methods 0.000 claims abstract description 11
- 238000004113 cell culture Methods 0.000 claims abstract description 8
- 238000005070 sampling Methods 0.000 claims abstract description 6
- 230000001131 transforming effect Effects 0.000 claims abstract description 6
- 235000018102 proteins Nutrition 0.000 claims description 209
- 108090000765 processed proteins & peptides Proteins 0.000 claims description 102
- 102000004190 Enzymes Human genes 0.000 claims description 42
- 108090000790 Enzymes Proteins 0.000 claims description 42
- 150000001413 amino acids Chemical class 0.000 claims description 39
- 235000001014 amino acid Nutrition 0.000 claims description 34
- 239000000427 antigen Substances 0.000 claims description 29
- 108091007433 antigens Proteins 0.000 claims description 28
- 102000036639 antigens Human genes 0.000 claims description 28
- 238000004458 analytical method Methods 0.000 claims description 25
- 230000000694 effects Effects 0.000 claims description 21
- 241001465754 Metazoa Species 0.000 claims description 19
- 229920000642 polymer Polymers 0.000 claims description 15
- 230000021615 conjugation Effects 0.000 claims description 13
- 239000007790 solid phase Substances 0.000 claims description 13
- 238000006467 substitution reaction Methods 0.000 claims description 13
- 241000282414 Homo sapiens Species 0.000 claims description 9
- 230000002860 competitive effect Effects 0.000 claims description 9
- 238000010790 dilution Methods 0.000 claims description 9
- 239000012895 dilution Substances 0.000 claims description 9
- 239000011324 bead Substances 0.000 claims description 8
- 235000018977 lysine Nutrition 0.000 claims description 8
- 230000004481 post-translational protein modification Effects 0.000 claims description 8
- 230000015572 biosynthetic process Effects 0.000 claims description 7
- 210000002966 serum Anatomy 0.000 claims description 7
- 239000007787 solid Substances 0.000 claims description 7
- 239000000126 substance Substances 0.000 claims description 7
- 108020004705 Codon Proteins 0.000 claims description 6
- 230000002411 adverse Effects 0.000 claims description 6
- 230000008859 change Effects 0.000 claims description 6
- 230000006870 function Effects 0.000 claims description 6
- 125000000539 amino acid group Chemical group 0.000 claims description 5
- 238000007385 chemical modification Methods 0.000 claims description 5
- 230000004927 fusion Effects 0.000 claims description 5
- 239000012535 impurity Substances 0.000 claims description 5
- 238000001727 in vivo Methods 0.000 claims description 5
- 150000002669 lysines Chemical class 0.000 claims description 5
- 230000026731 phosphorylation Effects 0.000 claims description 5
- 238000006366 phosphorylation reaction Methods 0.000 claims description 5
- 230000004988 N-glycosylation Effects 0.000 claims description 4
- 238000012217 deletion Methods 0.000 claims description 4
- 230000037430 deletion Effects 0.000 claims description 4
- 230000008569 process Effects 0.000 claims description 4
- 102000018697 Membrane Proteins Human genes 0.000 claims description 3
- 108010052285 Membrane Proteins Proteins 0.000 claims description 3
- 239000012228 culture supernatant Substances 0.000 claims description 3
- 239000012528 membrane Substances 0.000 claims description 3
- 230000012743 protein tagging Effects 0.000 claims description 3
- 239000012636 effector Substances 0.000 claims description 2
- 239000012085 test solution Substances 0.000 claims 4
- 230000004520 agglutination Effects 0.000 claims 1
- 230000009830 antibody antigen interaction Effects 0.000 claims 1
- 239000012560 cell impurity Substances 0.000 claims 1
- 210000004027 cell Anatomy 0.000 description 217
- 102000004196 processed proteins & peptides Human genes 0.000 description 84
- 229920001184 polypeptide Polymers 0.000 description 70
- 150000007523 nucleic acids Chemical class 0.000 description 69
- 240000004808 Saccharomyces cerevisiae Species 0.000 description 63
- 235000014680 Saccharomyces cerevisiae Nutrition 0.000 description 62
- 108091028043 Nucleic acid sequence Proteins 0.000 description 50
- 239000013598 vector Substances 0.000 description 43
- 108010076504 Protein Sorting Signals Proteins 0.000 description 41
- 229940088598 enzyme Drugs 0.000 description 41
- 108091026890 Coding region Proteins 0.000 description 39
- 108020004414 DNA Proteins 0.000 description 38
- 108091005804 Peptidases Proteins 0.000 description 37
- 230000002538 fungal effect Effects 0.000 description 37
- 239000004365 Protease Substances 0.000 description 35
- 102000035195 Peptidases Human genes 0.000 description 34
- 229940024606 amino acid Drugs 0.000 description 31
- 229920001223 polyethylene glycol Polymers 0.000 description 23
- 102000039446 nucleic acids Human genes 0.000 description 22
- 108020004707 nucleic acids Proteins 0.000 description 22
- 235000019419 proteases Nutrition 0.000 description 22
- 241000228245 Aspergillus niger Species 0.000 description 20
- 108090001060 Lipase Proteins 0.000 description 19
- 150000001875 compounds Chemical class 0.000 description 19
- 102000004882 Lipase Human genes 0.000 description 18
- 239000004367 Lipase Substances 0.000 description 17
- 235000019421 lipase Nutrition 0.000 description 17
- 241000233866 Fungi Species 0.000 description 16
- 240000006439 Aspergillus oryzae Species 0.000 description 15
- 235000014469 Bacillus subtilis Nutrition 0.000 description 15
- 241000228212 Aspergillus Species 0.000 description 14
- 238000006243 chemical reaction Methods 0.000 description 14
- 230000035772 mutation Effects 0.000 description 14
- 239000004382 Amylase Substances 0.000 description 13
- 108010065511 Amylases Proteins 0.000 description 13
- 235000002247 Aspergillus oryzae Nutrition 0.000 description 13
- 244000063299 Bacillus subtilis Species 0.000 description 13
- 102000004357 Transferases Human genes 0.000 description 13
- 108090000992 Transferases Proteins 0.000 description 13
- 208000026935 allergic disease Diseases 0.000 description 13
- -1 polyethylene Polymers 0.000 description 13
- 239000000523 sample Substances 0.000 description 13
- 238000012360 testing method Methods 0.000 description 13
- 238000013518 transcription Methods 0.000 description 13
- 230000035897 transcription Effects 0.000 description 13
- 230000005847 immunogenicity Effects 0.000 description 12
- 230000010076 replication Effects 0.000 description 12
- 241000193830 Bacillus <bacterium> Species 0.000 description 11
- 102000004316 Oxidoreductases Human genes 0.000 description 11
- 108090000854 Oxidoreductases Proteins 0.000 description 11
- 230000004913 activation Effects 0.000 description 11
- 239000002299 complementary DNA Substances 0.000 description 11
- 238000010168 coupling process Methods 0.000 description 11
- 239000006228 supernatant Substances 0.000 description 11
- 241000223218 Fusarium Species 0.000 description 10
- 230000008878 coupling Effects 0.000 description 10
- 238000005859 coupling reaction Methods 0.000 description 10
- 238000004519 manufacturing process Methods 0.000 description 10
- 230000008488 polyadenylation Effects 0.000 description 10
- 108010011619 6-Phytase Proteins 0.000 description 9
- 102000013142 Amylases Human genes 0.000 description 9
- 241000351920 Aspergillus nidulans Species 0.000 description 9
- 108060008539 Transglutaminase Proteins 0.000 description 9
- 235000019418 amylase Nutrition 0.000 description 9
- 230000001580 bacterial effect Effects 0.000 description 9
- 108010089934 carbohydrase Proteins 0.000 description 9
- RAXXELZNTBOGNW-UHFFFAOYSA-N imidazole Natural products C1=CNC=N1 RAXXELZNTBOGNW-UHFFFAOYSA-N 0.000 description 9
- 108010020132 microbial serine proteinases Proteins 0.000 description 9
- 230000001105 regulatory effect Effects 0.000 description 9
- 230000004044 response Effects 0.000 description 9
- 230000003248 secreting effect Effects 0.000 description 9
- 102000003601 transglutaminase Human genes 0.000 description 9
- 239000000872 buffer Substances 0.000 description 8
- 210000000349 chromosome Anatomy 0.000 description 8
- 239000013604 expression vector Substances 0.000 description 8
- 238000002744 homologous recombination Methods 0.000 description 8
- 230000006801 homologous recombination Effects 0.000 description 8
- 230000010354 integration Effects 0.000 description 8
- 239000000203 mixture Substances 0.000 description 8
- 238000002703 mutagenesis Methods 0.000 description 8
- 231100000350 mutagenesis Toxicity 0.000 description 8
- 239000012071 phase Substances 0.000 description 8
- 239000013612 plasmid Substances 0.000 description 8
- 230000009466 transformation Effects 0.000 description 8
- UHPMCKVQTMMPCG-UHFFFAOYSA-N 5,8-dihydroxy-2-methoxy-6-methyl-7-(2-oxopropyl)naphthalene-1,4-dione Chemical compound CC1=C(CC(C)=O)C(O)=C2C(=O)C(OC)=CC(=O)C2=C1O UHPMCKVQTMMPCG-UHFFFAOYSA-N 0.000 description 7
- 241000194108 Bacillus licheniformis Species 0.000 description 7
- 241000193385 Geobacillus stearothermophilus Species 0.000 description 7
- 108091034117 Oligonucleotide Proteins 0.000 description 7
- 230000002255 enzymatic effect Effects 0.000 description 7
- 230000028993 immune response Effects 0.000 description 7
- 239000000243 solution Substances 0.000 description 7
- 230000014616 translation Effects 0.000 description 7
- 108091032973 (ribonucleotides)n+m Proteins 0.000 description 6
- CSCPPACGZOOCGX-UHFFFAOYSA-N Acetone Chemical compound CC(C)=O CSCPPACGZOOCGX-UHFFFAOYSA-N 0.000 description 6
- 241000894006 Bacteria Species 0.000 description 6
- 241000222120 Candida <Saccharomycetales> Species 0.000 description 6
- 108091035707 Consensus sequence Proteins 0.000 description 6
- 102000053602 DNA Human genes 0.000 description 6
- 241000588724 Escherichia coli Species 0.000 description 6
- 229920001503 Glucan Polymers 0.000 description 6
- 206010020751 Hypersensitivity Diseases 0.000 description 6
- 241000228143 Penicillium Species 0.000 description 6
- 102000005158 Subtilisins Human genes 0.000 description 6
- 108010056079 Subtilisins Proteins 0.000 description 6
- JLCPHMBAVCMARE-UHFFFAOYSA-N [3-[[3-[[3-[[3-[[3-[[3-[[3-[[3-[[3-[[3-[[3-[[5-(2-amino-6-oxo-1H-purin-9-yl)-3-[[3-[[3-[[3-[[3-[[3-[[5-(2-amino-6-oxo-1H-purin-9-yl)-3-[[5-(2-amino-6-oxo-1H-purin-9-yl)-3-hydroxyoxolan-2-yl]methoxy-hydroxyphosphoryl]oxyoxolan-2-yl]methoxy-hydroxyphosphoryl]oxy-5-(5-methyl-2,4-dioxopyrimidin-1-yl)oxolan-2-yl]methoxy-hydroxyphosphoryl]oxy-5-(6-aminopurin-9-yl)oxolan-2-yl]methoxy-hydroxyphosphoryl]oxy-5-(6-aminopurin-9-yl)oxolan-2-yl]methoxy-hydroxyphosphoryl]oxy-5-(6-aminopurin-9-yl)oxolan-2-yl]methoxy-hydroxyphosphoryl]oxy-5-(6-aminopurin-9-yl)oxolan-2-yl]methoxy-hydroxyphosphoryl]oxyoxolan-2-yl]methoxy-hydroxyphosphoryl]oxy-5-(5-methyl-2,4-dioxopyrimidin-1-yl)oxolan-2-yl]methoxy-hydroxyphosphoryl]oxy-5-(4-amino-2-oxopyrimidin-1-yl)oxolan-2-yl]methoxy-hydroxyphosphoryl]oxy-5-(5-methyl-2,4-dioxopyrimidin-1-yl)oxolan-2-yl]methoxy-hydroxyphosphoryl]oxy-5-(5-methyl-2,4-dioxopyrimidin-1-yl)oxolan-2-yl]methoxy-hydroxyphosphoryl]oxy-5-(6-aminopurin-9-yl)oxolan-2-yl]methoxy-hydroxyphosphoryl]oxy-5-(6-aminopurin-9-yl)oxolan-2-yl]methoxy-hydroxyphosphoryl]oxy-5-(4-amino-2-oxopyrimidin-1-yl)oxolan-2-yl]methoxy-hydroxyphosphoryl]oxy-5-(4-amino-2-oxopyrimidin-1-yl)oxolan-2-yl]methoxy-hydroxyphosphoryl]oxy-5-(4-amino-2-oxopyrimidin-1-yl)oxolan-2-yl]methoxy-hydroxyphosphoryl]oxy-5-(6-aminopurin-9-yl)oxolan-2-yl]methoxy-hydroxyphosphoryl]oxy-5-(4-amino-2-oxopyrimidin-1-yl)oxolan-2-yl]methyl [5-(6-aminopurin-9-yl)-2-(hydroxymethyl)oxolan-3-yl] hydrogen phosphate Polymers Cc1cn(C2CC(OP(O)(=O)OCC3OC(CC3OP(O)(=O)OCC3OC(CC3O)n3cnc4c3nc(N)[nH]c4=O)n3cnc4c3nc(N)[nH]c4=O)C(COP(O)(=O)OC3CC(OC3COP(O)(=O)OC3CC(OC3COP(O)(=O)OC3CC(OC3COP(O)(=O)OC3CC(OC3COP(O)(=O)OC3CC(OC3COP(O)(=O)OC3CC(OC3COP(O)(=O)OC3CC(OC3COP(O)(=O)OC3CC(OC3COP(O)(=O)OC3CC(OC3COP(O)(=O)OC3CC(OC3COP(O)(=O)OC3CC(OC3COP(O)(=O)OC3CC(OC3COP(O)(=O)OC3CC(OC3COP(O)(=O)OC3CC(OC3COP(O)(=O)OC3CC(OC3COP(O)(=O)OC3CC(OC3COP(O)(=O)OC3CC(OC3CO)n3cnc4c(N)ncnc34)n3ccc(N)nc3=O)n3cnc4c(N)ncnc34)n3ccc(N)nc3=O)n3ccc(N)nc3=O)n3ccc(N)nc3=O)n3cnc4c(N)ncnc34)n3cnc4c(N)ncnc34)n3cc(C)c(=O)[nH]c3=O)n3cc(C)c(=O)[nH]c3=O)n3ccc(N)nc3=O)n3cc(C)c(=O)[nH]c3=O)n3cnc4c3nc(N)[nH]c4=O)n3cnc4c(N)ncnc34)n3cnc4c(N)ncnc34)n3cnc4c(N)ncnc34)n3cnc4c(N)ncnc34)O2)c(=O)[nH]c1=O JLCPHMBAVCMARE-UHFFFAOYSA-N 0.000 description 6
- 239000002253 acid Substances 0.000 description 6
- 230000007815 allergy Effects 0.000 description 6
- 210000004962 mammalian cell Anatomy 0.000 description 6
- 239000002609 medium Substances 0.000 description 6
- 244000005700 microbiome Species 0.000 description 6
- 238000003752 polymerase chain reaction Methods 0.000 description 6
- 238000013519 translation Methods 0.000 description 6
- 241000238631 Hexapoda Species 0.000 description 5
- 108090000787 Subtilisin Proteins 0.000 description 5
- 241000223259 Trichoderma Species 0.000 description 5
- 239000013566 allergen Substances 0.000 description 5
- 108090000637 alpha-Amylases Proteins 0.000 description 5
- 125000003277 amino group Chemical group 0.000 description 5
- 229940025131 amylases Drugs 0.000 description 5
- 230000008901 benefit Effects 0.000 description 5
- 239000001913 cellulose Substances 0.000 description 5
- 229920002678 cellulose Polymers 0.000 description 5
- 238000013461 design Methods 0.000 description 5
- 239000012634 fragment Substances 0.000 description 5
- 239000003550 marker Substances 0.000 description 5
- 239000000463 material Substances 0.000 description 5
- 108020004999 messenger RNA Proteins 0.000 description 5
- 230000004048 modification Effects 0.000 description 5
- 238000012986 modification Methods 0.000 description 5
- 230000007935 neutral effect Effects 0.000 description 5
- 230000037361 pathway Effects 0.000 description 5
- 239000013615 primer Substances 0.000 description 5
- 239000000047 product Substances 0.000 description 5
- 238000000746 purification Methods 0.000 description 5
- 230000028327 secretion Effects 0.000 description 5
- 239000000758 substrate Substances 0.000 description 5
- 230000002103 transcriptional effect Effects 0.000 description 5
- MTCFGRXMJLQNBG-REOHCLBHSA-N (2S)-2-Amino-3-hydroxypropansäure Chemical compound OC[C@H](N)C(O)=O MTCFGRXMJLQNBG-REOHCLBHSA-N 0.000 description 4
- 101000757144 Aspergillus niger Glucoamylase Proteins 0.000 description 4
- 108010084185 Cellulases Proteins 0.000 description 4
- 102000005575 Cellulases Human genes 0.000 description 4
- IAZDPXIOMUYVGZ-UHFFFAOYSA-N Dimethylsulphoxide Chemical compound CS(C)=O IAZDPXIOMUYVGZ-UHFFFAOYSA-N 0.000 description 4
- 241000223221 Fusarium oxysporum Species 0.000 description 4
- 108010073178 Glucan 1,4-alpha-Glucosidase Proteins 0.000 description 4
- 102100022624 Glucoamylase Human genes 0.000 description 4
- 241001480714 Humicola insolens Species 0.000 description 4
- 241000829100 Macaca mulatta polyomavirus 1 Species 0.000 description 4
- 241000235395 Mucor Species 0.000 description 4
- 241000226677 Myceliophthora Species 0.000 description 4
- 241000221960 Neurospora Species 0.000 description 4
- 102000003992 Peroxidases Human genes 0.000 description 4
- 241000235648 Pichia Species 0.000 description 4
- 108030003943 Protein-disulfide reductases Proteins 0.000 description 4
- 241000235403 Rhizomucor miehei Species 0.000 description 4
- 206010070834 Sensitisation Diseases 0.000 description 4
- MTCFGRXMJLQNBG-UHFFFAOYSA-N Serine Natural products OCC(N)C(O)=O MTCFGRXMJLQNBG-UHFFFAOYSA-N 0.000 description 4
- 210000001744 T-lymphocyte Anatomy 0.000 description 4
- 108010048241 acetamidase Proteins 0.000 description 4
- 125000002252 acyl group Chemical group 0.000 description 4
- 238000007792 addition Methods 0.000 description 4
- OIRDTQYFTABQOQ-KQYNXXCUSA-N adenosine Chemical compound C1=NC=2C(N)=NC=NC=2N1[C@@H]1O[C@H](CO)[C@@H](O)[C@H]1O OIRDTQYFTABQOQ-KQYNXXCUSA-N 0.000 description 4
- 150000001408 amides Chemical class 0.000 description 4
- 238000013459 approach Methods 0.000 description 4
- 210000003719 b-lymphocyte Anatomy 0.000 description 4
- 230000002759 chromosomal effect Effects 0.000 description 4
- 238000010367 cloning Methods 0.000 description 4
- 238000005516 engineering process Methods 0.000 description 4
- 238000010353 genetic engineering Methods 0.000 description 4
- 239000001963 growth medium Substances 0.000 description 4
- 230000007062 hydrolysis Effects 0.000 description 4
- 238000006460 hydrolysis reaction Methods 0.000 description 4
- 238000011534 incubation Methods 0.000 description 4
- 238000003780 insertion Methods 0.000 description 4
- 230000037431 insertion Effects 0.000 description 4
- 230000003993 interaction Effects 0.000 description 4
- 239000003446 ligand Substances 0.000 description 4
- 230000000813 microbial effect Effects 0.000 description 4
- 229920002704 polyhistidine Polymers 0.000 description 4
- 108091033319 polynucleotide Proteins 0.000 description 4
- 102000040430 polynucleotide Human genes 0.000 description 4
- 239000002157 polynucleotide Substances 0.000 description 4
- 229920000136 polysorbate Polymers 0.000 description 4
- 238000002360 preparation method Methods 0.000 description 4
- 125000002924 primary amino group Chemical group [H]N([H])* 0.000 description 4
- 238000012545 processing Methods 0.000 description 4
- 238000003259 recombinant expression Methods 0.000 description 4
- 235000020183 skimmed milk Nutrition 0.000 description 4
- 241000894007 species Species 0.000 description 4
- 108010037870 Anthranilate Synthase Proteins 0.000 description 3
- 241000193744 Bacillus amyloliquefaciens Species 0.000 description 3
- 101000950981 Bacillus subtilis (strain 168) Catabolic NAD-specific glutamate dehydrogenase RocG Proteins 0.000 description 3
- KXDHJXZQYSOELW-UHFFFAOYSA-M Carbamate Chemical compound NC([O-])=O KXDHJXZQYSOELW-UHFFFAOYSA-M 0.000 description 3
- 108010059892 Cellulase Proteins 0.000 description 3
- 238000002965 ELISA Methods 0.000 description 3
- 108010020195 FLAG peptide Proteins 0.000 description 3
- 241000221779 Fusarium sambucinum Species 0.000 description 3
- 108700007698 Genetic Terminator Regions Proteins 0.000 description 3
- 102000016901 Glutamate dehydrogenase Human genes 0.000 description 3
- 241000223198 Humicola Species 0.000 description 3
- 108010058683 Immobilized Proteins Proteins 0.000 description 3
- 108060003951 Immunoglobulin Proteins 0.000 description 3
- KDXKERNSBIXSRK-YFKPBYRVSA-N L-lysine Chemical compound NCCCC[C@H](N)C(O)=O KDXKERNSBIXSRK-YFKPBYRVSA-N 0.000 description 3
- 108010029541 Laccase Proteins 0.000 description 3
- 239000004472 Lysine Substances 0.000 description 3
- KDXKERNSBIXSRK-UHFFFAOYSA-N Lysine Natural products NCCCCC(N)C(O)=O KDXKERNSBIXSRK-UHFFFAOYSA-N 0.000 description 3
- 108010006519 Molecular Chaperones Proteins 0.000 description 3
- 241000233654 Oomycetes Species 0.000 description 3
- 241000283973 Oryctolagus cuniculus Species 0.000 description 3
- 102000012288 Phosphopyruvate Hydratase Human genes 0.000 description 3
- 108010022181 Phosphopyruvate Hydratase Proteins 0.000 description 3
- 108091000080 Phosphotransferase Proteins 0.000 description 3
- 229920001213 Polysorbate 20 Polymers 0.000 description 3
- 108020004511 Recombinant DNA Proteins 0.000 description 3
- 101000968489 Rhizomucor miehei Lipase Proteins 0.000 description 3
- 241000223258 Thermomyces lanuginosus Species 0.000 description 3
- 241001494489 Thielavia Species 0.000 description 3
- 241000499912 Trichoderma reesei Species 0.000 description 3
- 108700040099 Xylose isomerases Proteins 0.000 description 3
- 239000012190 activator Substances 0.000 description 3
- 102000004139 alpha-Amylases Human genes 0.000 description 3
- 229940024171 alpha-amylase Drugs 0.000 description 3
- 230000003321 amplification Effects 0.000 description 3
- 238000010171 animal model Methods 0.000 description 3
- 210000003651 basophil Anatomy 0.000 description 3
- PFYXSUNOLOJMDX-UHFFFAOYSA-N bis(2,5-dioxopyrrolidin-1-yl) carbonate Chemical compound O=C1CCC(=O)N1OC(=O)ON1C(=O)CCC1=O PFYXSUNOLOJMDX-UHFFFAOYSA-N 0.000 description 3
- 230000000903 blocking effect Effects 0.000 description 3
- 230000001413 cellular effect Effects 0.000 description 3
- 238000004587 chromatography analysis Methods 0.000 description 3
- MGNCLNQXLYJVJD-UHFFFAOYSA-N cyanuric chloride Chemical compound ClC1=NC(Cl)=NC(Cl)=N1 MGNCLNQXLYJVJD-UHFFFAOYSA-N 0.000 description 3
- 238000001514 detection method Methods 0.000 description 3
- 238000009826 distribution Methods 0.000 description 3
- AFOSIXZFDONLBT-UHFFFAOYSA-N divinyl sulfone Chemical compound C=CS(=O)(=O)C=C AFOSIXZFDONLBT-UHFFFAOYSA-N 0.000 description 3
- 108010061330 glucan 1,4-alpha-maltohydrolase Proteins 0.000 description 3
- 210000003630 histaminocyte Anatomy 0.000 description 3
- 230000002209 hydrophobic effect Effects 0.000 description 3
- 125000002887 hydroxy group Chemical group [H]O* 0.000 description 3
- 102000018358 immunoglobulin Human genes 0.000 description 3
- 238000000338 in vitro Methods 0.000 description 3
- 125000003588 lysine group Chemical group [H]N([H])C([H])([H])C([H])([H])C([H])([H])C([H])([H])C([H])(N([H])[H])C(*)=O 0.000 description 3
- 238000003199 nucleic acid amplification method Methods 0.000 description 3
- 239000002773 nucleotide Substances 0.000 description 3
- 125000003729 nucleotide group Chemical group 0.000 description 3
- 108040007629 peroxidase activity proteins Proteins 0.000 description 3
- 239000000256 polyoxyethylene sorbitan monolaurate Substances 0.000 description 3
- 235000010486 polyoxyethylene sorbitan monolaurate Nutrition 0.000 description 3
- 239000000843 powder Substances 0.000 description 3
- 238000000159 protein binding assay Methods 0.000 description 3
- 210000001938 protoplast Anatomy 0.000 description 3
- 230000006798 recombination Effects 0.000 description 3
- 238000005215 recombination Methods 0.000 description 3
- 230000009467 reduction Effects 0.000 description 3
- 150000003839 salts Chemical class 0.000 description 3
- 210000003491 skin Anatomy 0.000 description 3
- 239000011343 solid material Substances 0.000 description 3
- 238000010561 standard procedure Methods 0.000 description 3
- 238000003786 synthesis reaction Methods 0.000 description 3
- 239000004753 textile Substances 0.000 description 3
- 210000005253 yeast cell Anatomy 0.000 description 3
- LZKGFGLOQNSMBS-UHFFFAOYSA-N 4,5,6-trichlorotriazine Chemical compound ClC1=NN=NC(Cl)=C1Cl LZKGFGLOQNSMBS-UHFFFAOYSA-N 0.000 description 2
- FHVDTGUDJYJELY-UHFFFAOYSA-N 6-{[2-carboxy-4,5-dihydroxy-6-(phosphanyloxy)oxan-3-yl]oxy}-4,5-dihydroxy-3-phosphanyloxane-2-carboxylic acid Chemical compound O1C(C(O)=O)C(P)C(O)C(O)C1OC1C(C(O)=O)OC(OP)C(O)C1O FHVDTGUDJYJELY-UHFFFAOYSA-N 0.000 description 2
- 241001019659 Acremonium <Plectosphaerellaceae> Species 0.000 description 2
- 108010021809 Alcohol dehydrogenase Proteins 0.000 description 2
- 102100034042 Alcohol dehydrogenase 1C Human genes 0.000 description 2
- 102100034044 All-trans-retinol dehydrogenase [NAD(+)] ADH1B Human genes 0.000 description 2
- 101710193111 All-trans-retinol dehydrogenase [NAD(+)] ADH4 Proteins 0.000 description 2
- 108090000915 Aminopeptidases Proteins 0.000 description 2
- 102000004400 Aminopeptidases Human genes 0.000 description 2
- QGZKDVFQNNGYKY-UHFFFAOYSA-N Ammonia Chemical compound N QGZKDVFQNNGYKY-UHFFFAOYSA-N 0.000 description 2
- 241000235349 Ascomycota Species 0.000 description 2
- 108010017640 Aspartic Acid Proteases Proteins 0.000 description 2
- 241001513093 Aspergillus awamori Species 0.000 description 2
- 101000690713 Aspergillus niger Alpha-glucosidase Proteins 0.000 description 2
- 108090000145 Bacillolysin Proteins 0.000 description 2
- 241000193422 Bacillus lentus Species 0.000 description 2
- 101000695691 Bacillus licheniformis Beta-lactamase Proteins 0.000 description 2
- 241000194103 Bacillus pumilus Species 0.000 description 2
- 108091005658 Basic proteases Proteins 0.000 description 2
- 241000221198 Basidiomycota Species 0.000 description 2
- 241000283690 Bos taurus Species 0.000 description 2
- 241000701822 Bovine papillomavirus Species 0.000 description 2
- OKTJSMMVPCPJKN-UHFFFAOYSA-N Carbon Chemical compound [C] OKTJSMMVPCPJKN-UHFFFAOYSA-N 0.000 description 2
- 102000004308 Carboxylic Ester Hydrolases Human genes 0.000 description 2
- 108090000863 Carboxylic Ester Hydrolases Proteins 0.000 description 2
- 241000700199 Cavia porcellus Species 0.000 description 2
- 229920001661 Chitosan Polymers 0.000 description 2
- 241000233652 Chytridiomycota Species 0.000 description 2
- 244000251987 Coprinus macrorhizus Species 0.000 description 2
- 235000001673 Coprinus macrorhizus Nutrition 0.000 description 2
- 101000796894 Coturnix japonica Alcohol dehydrogenase 1 Proteins 0.000 description 2
- 108010003989 D-amino-acid oxidase Proteins 0.000 description 2
- 102000004674 D-amino-acid oxidase Human genes 0.000 description 2
- SRBFZHDQGSBBOR-IOVATXLUSA-N D-xylopyranose Chemical compound O[C@@H]1COC(O)[C@H](O)[C@H]1O SRBFZHDQGSBBOR-IOVATXLUSA-N 0.000 description 2
- 102000004163 DNA-directed RNA polymerases Human genes 0.000 description 2
- 108090000626 DNA-directed RNA polymerases Proteins 0.000 description 2
- 108010001682 Dextranase Proteins 0.000 description 2
- 101710121765 Endo-1,4-beta-xylanase Proteins 0.000 description 2
- 102000010911 Enzyme Precursors Human genes 0.000 description 2
- 108010062466 Enzyme Precursors Proteins 0.000 description 2
- 241000206602 Eukaryota Species 0.000 description 2
- XZWYTXMRWQJBGX-VXBMVYAYSA-N FLAG peptide Chemical compound NCCCC[C@@H](C(O)=O)NC(=O)[C@H](CC(O)=O)NC(=O)[C@H](CC(O)=O)NC(=O)[C@H](CC(O)=O)NC(=O)[C@H](CC(O)=O)NC(=O)[C@H](CCCCN)NC(=O)[C@@H](NC(=O)[C@@H](N)CC(O)=O)CC1=CC=C(O)C=C1 XZWYTXMRWQJBGX-VXBMVYAYSA-N 0.000 description 2
- 241000567163 Fusarium cerealis Species 0.000 description 2
- 241000146406 Fusarium heterosporum Species 0.000 description 2
- 241000427940 Fusarium solani Species 0.000 description 2
- 101150108358 GLAA gene Proteins 0.000 description 2
- 101100369308 Geobacillus stearothermophilus nprS gene Proteins 0.000 description 2
- 101100080316 Geobacillus stearothermophilus nprT gene Proteins 0.000 description 2
- 108010015776 Glucose oxidase Proteins 0.000 description 2
- 102000004366 Glucosidases Human genes 0.000 description 2
- 108010056771 Glucosidases Proteins 0.000 description 2
- WHUUTDBJXJRKMK-UHFFFAOYSA-N Glutamic acid Natural products OC(=O)C(N)CCC(O)=O WHUUTDBJXJRKMK-UHFFFAOYSA-N 0.000 description 2
- 102000000587 Glycerolphosphate Dehydrogenase Human genes 0.000 description 2
- 108010041921 Glycerolphosphate Dehydrogenase Proteins 0.000 description 2
- NYHBQMYGNKIUIF-UUOKFMHZSA-N Guanosine Chemical compound C1=NC=2C(=O)NC(N)=NC=2N1[C@@H]1O[C@H](CO)[C@@H](O)[C@H]1O NYHBQMYGNKIUIF-UUOKFMHZSA-N 0.000 description 2
- 101000780463 Homo sapiens Alcohol dehydrogenase 1C Proteins 0.000 description 2
- 101000801742 Homo sapiens Triosephosphate isomerase Proteins 0.000 description 2
- 102000009438 IgE Receptors Human genes 0.000 description 2
- 108010073816 IgE Receptors Proteins 0.000 description 2
- 102000004195 Isomerases Human genes 0.000 description 2
- 108090000769 Isomerases Proteins 0.000 description 2
- 102100027612 Kallikrein-11 Human genes 0.000 description 2
- 241000235649 Kluyveromyces Species 0.000 description 2
- 244000285963 Kluyveromyces fragilis Species 0.000 description 2
- 235000014663 Kluyveromyces fragilis Nutrition 0.000 description 2
- ZDXPYRJPNDTMRX-VKHMYHEASA-N L-glutamine Chemical compound OC(=O)[C@@H](N)CCC(N)=O ZDXPYRJPNDTMRX-VKHMYHEASA-N 0.000 description 2
- FBOZXECLQNJBKD-ZDUSSCGKSA-N L-methotrexate Chemical compound C=1N=C2N=C(N)N=C(N)C2=NC=1CN(C)C1=CC=C(C(=O)N[C@@H](CCC(O)=O)C(O)=O)C=C1 FBOZXECLQNJBKD-ZDUSSCGKSA-N 0.000 description 2
- OUYCCCASQSFEME-QMMMGPOBSA-N L-tyrosine Chemical compound OC(=O)[C@@H](N)CC1=CC=C(O)C=C1 OUYCCCASQSFEME-QMMMGPOBSA-N 0.000 description 2
- 108010028921 Lipopeptides Proteins 0.000 description 2
- 102000004317 Lyases Human genes 0.000 description 2
- 108090000856 Lyases Proteins 0.000 description 2
- 108090000157 Metallothionein Proteins 0.000 description 2
- AFVFQIVMOAPDHO-UHFFFAOYSA-N Methanesulfonic acid Chemical class CS(O)(=O)=O AFVFQIVMOAPDHO-UHFFFAOYSA-N 0.000 description 2
- NQTADLQHYWFPDB-UHFFFAOYSA-N N-Hydroxysuccinimide Chemical compound ON1C(=O)CCC1=O NQTADLQHYWFPDB-UHFFFAOYSA-N 0.000 description 2
- PXHVJJICTQNCMI-UHFFFAOYSA-N Nickel Chemical compound [Ni] PXHVJJICTQNCMI-UHFFFAOYSA-N 0.000 description 2
- 108700026244 Open Reading Frames Proteins 0.000 description 2
- 108010033276 Peptide Fragments Proteins 0.000 description 2
- 102000007079 Peptide Fragments Human genes 0.000 description 2
- 108010059820 Polygalacturonase Proteins 0.000 description 2
- 102000006010 Protein Disulfide-Isomerase Human genes 0.000 description 2
- 241000589516 Pseudomonas Species 0.000 description 2
- 241000813090 Rhizoctonia solani Species 0.000 description 2
- 241000235402 Rhizomucor Species 0.000 description 2
- 241000235527 Rhizopus Species 0.000 description 2
- 241000714474 Rous sarcoma virus Species 0.000 description 2
- 241000235070 Saccharomyces Species 0.000 description 2
- 241000235346 Schizosaccharomyces Species 0.000 description 2
- CDBYLPFSWZWCQE-UHFFFAOYSA-L Sodium Carbonate Chemical compound [Na+].[Na+].[O-]C([O-])=O CDBYLPFSWZWCQE-UHFFFAOYSA-L 0.000 description 2
- FAPWRFPIFSIZLT-UHFFFAOYSA-M Sodium chloride Chemical compound [Na+].[Cl-] FAPWRFPIFSIZLT-UHFFFAOYSA-M 0.000 description 2
- 108091081024 Start codon Proteins 0.000 description 2
- 241000509227 Stercorarius antarcticus Species 0.000 description 2
- 241000187747 Streptomyces Species 0.000 description 2
- 108010022394 Threonine synthase Proteins 0.000 description 2
- 102000036693 Thrombopoietin Human genes 0.000 description 2
- 108010041111 Thrombopoietin Proteins 0.000 description 2
- IQFYYKKMVGJFEH-XLPZGREQSA-N Thymidine Chemical compound O=C1NC(=O)C(C)=CN1[C@@H]1O[C@H](CO)[C@@H](O)C1 IQFYYKKMVGJFEH-XLPZGREQSA-N 0.000 description 2
- 241001149964 Tolypocladium Species 0.000 description 2
- 241000223260 Trichoderma harzianum Species 0.000 description 2
- 102100033598 Triosephosphate isomerase Human genes 0.000 description 2
- 108090000631 Trypsin Proteins 0.000 description 2
- 102000004142 Trypsin Human genes 0.000 description 2
- 101710152431 Trypsin-like protease Proteins 0.000 description 2
- DRTQHJPVMGBUCF-XVFCMESISA-N Uridine Chemical compound O[C@@H]1[C@H](O)[C@@H](CO)O[C@H]1N1C(=O)NC(=O)C=C1 DRTQHJPVMGBUCF-XVFCMESISA-N 0.000 description 2
- 241000700605 Viruses Species 0.000 description 2
- IXKSXJFAGXLQOQ-XISFHERQSA-N WHWLQLKPGQPMY Chemical compound C([C@@H](C(=O)N[C@@H](CC=1C2=CC=CC=C2NC=1)C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](CCC(N)=O)C(=O)N[C@@H](CC(C)C)C(=O)N1CCC[C@H]1C(=O)NCC(=O)N[C@@H](CCC(N)=O)C(=O)N[C@@H](CC(O)=O)C(=O)N1CCC[C@H]1C(=O)N[C@@H](CCSC)C(=O)N[C@@H](CC=1C=CC(O)=CC=1)C(O)=O)NC(=O)[C@@H](N)CC=1C2=CC=CC=C2NC=1)C1=CNC=N1 IXKSXJFAGXLQOQ-XISFHERQSA-N 0.000 description 2
- 241000758405 Zoopagomycotina Species 0.000 description 2
- XJLXINKUBYWONI-DQQFMEOOSA-N [[(2r,3r,4r,5r)-5-(6-aminopurin-9-yl)-3-hydroxy-4-phosphonooxyoxolan-2-yl]methoxy-hydroxyphosphoryl] [(2s,3r,4s,5s)-5-(3-carbamoylpyridin-1-ium-1-yl)-3,4-dihydroxyoxolan-2-yl]methyl phosphate Chemical compound NC(=O)C1=CC=C[N+]([C@@H]2[C@H]([C@@H](O)[C@H](COP([O-])(=O)OP(O)(=O)OC[C@@H]3[C@H]([C@@H](OP(O)(O)=O)[C@@H](O3)N3C4=NC=NC(N)=C4N=C3)O)O2)O)=C1 XJLXINKUBYWONI-DQQFMEOOSA-N 0.000 description 2
- 239000000370 acceptor Substances 0.000 description 2
- 230000002378 acidificating effect Effects 0.000 description 2
- 150000001299 aldehydes Chemical class 0.000 description 2
- 229940072056 alginate Drugs 0.000 description 2
- 235000010443 alginic acid Nutrition 0.000 description 2
- 229920000615 alginic acid Polymers 0.000 description 2
- 230000002009 allergenic effect Effects 0.000 description 2
- 230000000845 anti-microbial effect Effects 0.000 description 2
- 238000011091 antibody purification Methods 0.000 description 2
- 239000004599 antimicrobial Substances 0.000 description 2
- 125000000637 arginyl group Chemical class N[C@@H](CCCNC(N)=N)C(=O)* 0.000 description 2
- 230000003115 biocidal effect Effects 0.000 description 2
- 230000033228 biological regulation Effects 0.000 description 2
- 150000001718 carbodiimides Chemical class 0.000 description 2
- 150000001720 carbohydrates Chemical group 0.000 description 2
- 229910052799 carbon Inorganic materials 0.000 description 2
- 125000003178 carboxy group Chemical group [H]OC(*)=O 0.000 description 2
- 239000005018 casein Substances 0.000 description 2
- 230000015556 catabolic process Effects 0.000 description 2
- 210000002421 cell wall Anatomy 0.000 description 2
- 229940106157 cellulase Drugs 0.000 description 2
- 238000005119 centrifugation Methods 0.000 description 2
- 239000003795 chemical substances by application Substances 0.000 description 2
- FZFAMSAMCHXGEF-UHFFFAOYSA-N chloro formate Chemical compound ClOC=O FZFAMSAMCHXGEF-UHFFFAOYSA-N 0.000 description 2
- 238000004132 cross linking Methods 0.000 description 2
- 239000013078 crystal Substances 0.000 description 2
- 230000003247 decreasing effect Effects 0.000 description 2
- 230000001419 dependent effect Effects 0.000 description 2
- 239000003599 detergent Substances 0.000 description 2
- XBDQKXXYIPTUBI-UHFFFAOYSA-N dimethylselenoniopropionate Natural products CCC(O)=O XBDQKXXYIPTUBI-UHFFFAOYSA-N 0.000 description 2
- 239000003814 drug Substances 0.000 description 2
- 239000003623 enhancer Substances 0.000 description 2
- 239000002532 enzyme inhibitor Substances 0.000 description 2
- 210000003527 eukaryotic cell Anatomy 0.000 description 2
- 238000001914 filtration Methods 0.000 description 2
- 125000000524 functional group Chemical group 0.000 description 2
- 230000002068 genetic effect Effects 0.000 description 2
- 235000019420 glucose oxidase Nutrition 0.000 description 2
- 235000013922 glutamic acid Nutrition 0.000 description 2
- 239000004220 glutamic acid Substances 0.000 description 2
- ZDXPYRJPNDTMRX-UHFFFAOYSA-N glutamine Natural products OC(=O)C(N)CCC(N)=O ZDXPYRJPNDTMRX-UHFFFAOYSA-N 0.000 description 2
- 108020004445 glyceraldehyde-3-phosphate dehydrogenase Proteins 0.000 description 2
- 125000003147 glycosyl group Chemical group 0.000 description 2
- 230000012010 growth Effects 0.000 description 2
- 229910001385 heavy metal Inorganic materials 0.000 description 2
- 210000004408 hybridoma Anatomy 0.000 description 2
- 230000009851 immunogenic response Effects 0.000 description 2
- 238000010348 incorporation Methods 0.000 description 2
- NOESYZHRGYRDHS-UHFFFAOYSA-N insulin Chemical compound N1C(=O)C(NC(=O)C(CCC(N)=O)NC(=O)C(CCC(O)=O)NC(=O)C(C(C)C)NC(=O)C(NC(=O)CN)C(C)CC)CSSCC(C(NC(CO)C(=O)NC(CC(C)C)C(=O)NC(CC=2C=CC(O)=CC=2)C(=O)NC(CCC(N)=O)C(=O)NC(CC(C)C)C(=O)NC(CCC(O)=O)C(=O)NC(CC(N)=O)C(=O)NC(CC=2C=CC(O)=CC=2)C(=O)NC(CSSCC(NC(=O)C(C(C)C)NC(=O)C(CC(C)C)NC(=O)C(CC=2C=CC(O)=CC=2)NC(=O)C(CC(C)C)NC(=O)C(C)NC(=O)C(CCC(O)=O)NC(=O)C(C(C)C)NC(=O)C(CC(C)C)NC(=O)C(CC=2NC=NC=2)NC(=O)C(CO)NC(=O)CNC2=O)C(=O)NCC(=O)NC(CCC(O)=O)C(=O)NC(CCCNC(N)=N)C(=O)NCC(=O)NC(CC=3C=CC=CC=3)C(=O)NC(CC=3C=CC=CC=3)C(=O)NC(CC=3C=CC(O)=CC=3)C(=O)NC(C(C)O)C(=O)N3C(CCC3)C(=O)NC(CCCCN)C(=O)NC(C)C(O)=O)C(=O)NC(CC(N)=O)C(O)=O)=O)NC(=O)C(C(C)CC)NC(=O)C(CO)NC(=O)C(C(C)O)NC(=O)C1CSSCC2NC(=O)C(CC(C)C)NC(=O)C(NC(=O)C(CCC(N)=O)NC(=O)C(CC(N)=O)NC(=O)C(NC(=O)C(N)CC=1C=CC=CC=1)C(C)C)CC1=CN=CN1 NOESYZHRGYRDHS-UHFFFAOYSA-N 0.000 description 2
- 230000003834 intracellular effect Effects 0.000 description 2
- 238000001155 isoelectric focusing Methods 0.000 description 2
- 238000002955 isolation Methods 0.000 description 2
- 238000007834 ligase chain reaction Methods 0.000 description 2
- 230000000670 limiting effect Effects 0.000 description 2
- 238000005259 measurement Methods 0.000 description 2
- 108010003855 mesentericopeptidase Proteins 0.000 description 2
- 229960000485 methotrexate Drugs 0.000 description 2
- 230000002906 microbiologic effect Effects 0.000 description 2
- 238000010369 molecular cloning Methods 0.000 description 2
- 235000015097 nutrients Nutrition 0.000 description 2
- WWZKQHOCKIZLMA-UHFFFAOYSA-N octanoic acid Chemical compound CCCCCCCC(O)=O WWZKQHOCKIZLMA-UHFFFAOYSA-N 0.000 description 2
- 230000001717 pathogenic effect Effects 0.000 description 2
- KHIWWQKSHDUIBK-UHFFFAOYSA-N periodic acid Chemical compound OI(=O)(=O)=O KHIWWQKSHDUIBK-UHFFFAOYSA-N 0.000 description 2
- 238000002823 phage display Methods 0.000 description 2
- YBYRMVIVWMBXKQ-UHFFFAOYSA-N phenylmethanesulfonyl fluoride Chemical compound FS(=O)(=O)CC1=CC=CC=C1 YBYRMVIVWMBXKQ-UHFFFAOYSA-N 0.000 description 2
- 102000020233 phosphotransferase Human genes 0.000 description 2
- 229920001282 polysaccharide Polymers 0.000 description 2
- 239000005017 polysaccharide Substances 0.000 description 2
- 108091022901 polysaccharide lyase Proteins 0.000 description 2
- 102000020244 polysaccharide lyase Human genes 0.000 description 2
- 150000004804 polysaccharides Chemical class 0.000 description 2
- 108020003519 protein disulfide isomerase Proteins 0.000 description 2
- 101150054232 pyrG gene Proteins 0.000 description 2
- 108020003175 receptors Proteins 0.000 description 2
- 102000005962 receptors Human genes 0.000 description 2
- 238000011160 research Methods 0.000 description 2
- 230000008313 sensitization Effects 0.000 description 2
- 238000000926 separation method Methods 0.000 description 2
- 150000003461 sulfonyl halides Chemical class 0.000 description 2
- 125000003396 thiol group Chemical group [H]S* 0.000 description 2
- 210000001519 tissue Anatomy 0.000 description 2
- 239000012588 trypsin Substances 0.000 description 2
- OUYCCCASQSFEME-UHFFFAOYSA-N tyrosine Natural products OC(=O)C(N)CC1=CC=C(O)C=C1 OUYCCCASQSFEME-UHFFFAOYSA-N 0.000 description 2
- 241000701447 unidentified baculovirus Species 0.000 description 2
- 230000009105 vegetative growth Effects 0.000 description 2
- 238000005406 washing Methods 0.000 description 2
- CIAMOALLBVRDDK-UHFFFAOYSA-N 1,1-diaminopropan-2-ol Chemical compound CC(O)C(N)N CIAMOALLBVRDDK-UHFFFAOYSA-N 0.000 description 1
- GEYOCULIXLDCMW-UHFFFAOYSA-N 1,2-phenylenediamine Chemical compound NC1=CC=CC=C1N GEYOCULIXLDCMW-UHFFFAOYSA-N 0.000 description 1
- RYHBNJHYFVUHQT-UHFFFAOYSA-N 1,4-Dioxane Chemical compound C1COCCO1 RYHBNJHYFVUHQT-UHFFFAOYSA-N 0.000 description 1
- UHDGCWIWMRVCDJ-UHFFFAOYSA-N 1-beta-D-Xylofuranosyl-NH-Cytosine Natural products O=C1N=C(N)C=CN1C1C(O)C(O)C(CO)O1 UHDGCWIWMRVCDJ-UHFFFAOYSA-N 0.000 description 1
- OWEGMIWEEQEYGQ-UHFFFAOYSA-N 100676-05-9 Natural products OC1C(O)C(O)C(CO)OC1OCC1C(O)C(O)C(O)C(OC2C(OC(O)C(O)C2O)CO)O1 OWEGMIWEEQEYGQ-UHFFFAOYSA-N 0.000 description 1
- YKBGVTZYEHREMT-KVQBGUIXSA-N 2'-deoxyguanosine Chemical compound C1=NC=2C(=O)NC(N)=NC=2N1[C@H]1C[C@H](O)[C@@H](CO)O1 YKBGVTZYEHREMT-KVQBGUIXSA-N 0.000 description 1
- CKTSBUTUHBMZGZ-SHYZEUOFSA-N 2'‐deoxycytidine Chemical compound O=C1N=C(N)C=CN1[C@@H]1O[C@H](CO)[C@@H](O)C1 CKTSBUTUHBMZGZ-SHYZEUOFSA-N 0.000 description 1
- CXCHEKCRJQRVNG-UHFFFAOYSA-N 2,2,2-trifluoroethanesulfonyl chloride Chemical compound FC(F)(F)CS(Cl)(=O)=O CXCHEKCRJQRVNG-UHFFFAOYSA-N 0.000 description 1
- IEJPPSMHUUQABK-UHFFFAOYSA-N 2,4-diphenyl-4h-1,3-oxazol-5-one Chemical class O=C1OC(C=2C=CC=CC=2)=NC1C1=CC=CC=C1 IEJPPSMHUUQABK-UHFFFAOYSA-N 0.000 description 1
- FALRKNHUBBKYCC-UHFFFAOYSA-N 2-(chloromethyl)pyridine-3-carbonitrile Chemical compound ClCC1=NC=CC=C1C#N FALRKNHUBBKYCC-UHFFFAOYSA-N 0.000 description 1
- KISWVXRQTGLFGD-UHFFFAOYSA-N 2-[[2-[[6-amino-2-[[2-[[2-[[5-amino-2-[[2-[[1-[2-[[6-amino-2-[(2,5-diamino-5-oxopentanoyl)amino]hexanoyl]amino]-5-(diaminomethylideneamino)pentanoyl]pyrrolidine-2-carbonyl]amino]-3-hydroxypropanoyl]amino]-5-oxopentanoyl]amino]-5-(diaminomethylideneamino)p Chemical compound C1CCN(C(=O)C(CCCN=C(N)N)NC(=O)C(CCCCN)NC(=O)C(N)CCC(N)=O)C1C(=O)NC(CO)C(=O)NC(CCC(N)=O)C(=O)NC(CCCN=C(N)N)C(=O)NC(CO)C(=O)NC(CCCCN)C(=O)NC(C(=O)NC(CC(C)C)C(O)=O)CC1=CC=C(O)C=C1 KISWVXRQTGLFGD-UHFFFAOYSA-N 0.000 description 1
- QKNYBSVHEMOAJP-UHFFFAOYSA-N 2-amino-2-(hydroxymethyl)propane-1,3-diol;hydron;chloride Chemical compound Cl.OCC(N)(CO)CO QKNYBSVHEMOAJP-UHFFFAOYSA-N 0.000 description 1
- IAJOBQBIJHVGMQ-UHFFFAOYSA-N 2-amino-4-[hydroxy(methyl)phosphoryl]butanoic acid Chemical compound CP(O)(=O)CCC(N)C(O)=O IAJOBQBIJHVGMQ-UHFFFAOYSA-N 0.000 description 1
- BFSVOASYOCHEOV-UHFFFAOYSA-N 2-diethylaminoethanol Chemical compound CCN(CC)CCO BFSVOASYOCHEOV-UHFFFAOYSA-N 0.000 description 1
- OSJPPGNTCRNQQC-UWTATZPHSA-N 3-phospho-D-glyceric acid Chemical compound OC(=O)[C@H](O)COP(O)(O)=O OSJPPGNTCRNQQC-UWTATZPHSA-N 0.000 description 1
- 108010080981 3-phytase Proteins 0.000 description 1
- 238000010600 3H thymidine incorporation assay Methods 0.000 description 1
- LKDMKWNDBAVNQZ-WJNSRDFLSA-N 4-[[(2s)-1-[[(2s)-1-[(2s)-2-[[(2s)-1-(4-nitroanilino)-1-oxo-3-phenylpropan-2-yl]carbamoyl]pyrrolidin-1-yl]-1-oxopropan-2-yl]amino]-1-oxopropan-2-yl]amino]-4-oxobutanoic acid Chemical compound OC(=O)CCC(=O)N[C@@H](C)C(=O)N[C@@H](C)C(=O)N1CCC[C@H]1C(=O)N[C@H](C(=O)NC=1C=CC(=CC=1)[N+]([O-])=O)CC1=CC=CC=C1 LKDMKWNDBAVNQZ-WJNSRDFLSA-N 0.000 description 1
- JVVRCYWZTJLJSG-UHFFFAOYSA-N 4-dimethylaminophenol Chemical compound CN(C)C1=CC=C(O)C=C1 JVVRCYWZTJLJSG-UHFFFAOYSA-N 0.000 description 1
- 229960000549 4-dimethylaminophenol Drugs 0.000 description 1
- VHYFNPMBLIVWCW-UHFFFAOYSA-N 4-dimethylaminopyridine Substances CN(C)C1=CC=NC=C1 VHYFNPMBLIVWCW-UHFFFAOYSA-N 0.000 description 1
- TYMLOMAKGOJONV-UHFFFAOYSA-N 4-nitroaniline Chemical compound NC1=CC=C([N+]([O-])=O)C=C1 TYMLOMAKGOJONV-UHFFFAOYSA-N 0.000 description 1
- BTJIUGUIPKRLHP-UHFFFAOYSA-N 4-nitrophenol Chemical compound OC1=CC=C([N+]([O-])=O)C=C1 BTJIUGUIPKRLHP-UHFFFAOYSA-N 0.000 description 1
- 101710163881 5,6-dihydroxyindole-2-carboxylic acid oxidase Proteins 0.000 description 1
- RZVAJINKPMORJF-UHFFFAOYSA-N Acetaminophen Chemical compound CC(=O)NC1=CC=C(O)C=C1 RZVAJINKPMORJF-UHFFFAOYSA-N 0.000 description 1
- 241001578974 Achlya <moth> Species 0.000 description 1
- 241000228431 Acremonium chrysogenum Species 0.000 description 1
- 102000057234 Acyl transferases Human genes 0.000 description 1
- 108700016155 Acyl transferases Proteins 0.000 description 1
- 101800002326 Adipokinetic hormone Proteins 0.000 description 1
- 229920001817 Agar Polymers 0.000 description 1
- 235000001674 Agaricus brunnescens Nutrition 0.000 description 1
- 229920000936 Agarose Polymers 0.000 description 1
- 108010031025 Alanine Dehydrogenase Proteins 0.000 description 1
- 241000588986 Alcaligenes Species 0.000 description 1
- 244000300657 Alchornea rugosa Species 0.000 description 1
- 102000007698 Alcohol dehydrogenase Human genes 0.000 description 1
- 108020004774 Alkaline Phosphatase Proteins 0.000 description 1
- 102000002260 Alkaline Phosphatase Human genes 0.000 description 1
- 208000035285 Allergic Seasonal Rhinitis Diseases 0.000 description 1
- 241001136561 Allomyces Species 0.000 description 1
- 241000223600 Alternaria Species 0.000 description 1
- 108030000961 Aminopeptidase Y Proteins 0.000 description 1
- 108090000886 Ananain Proteins 0.000 description 1
- 206010002199 Anaphylactic shock Diseases 0.000 description 1
- 241000534414 Anotopterus nikparini Species 0.000 description 1
- 108700042778 Antimicrobial Peptides Proteins 0.000 description 1
- 102000044503 Antimicrobial Peptides Human genes 0.000 description 1
- 101100163849 Arabidopsis thaliana ARS1 gene Proteins 0.000 description 1
- 239000004475 Arginine Substances 0.000 description 1
- 240000003291 Armoracia rusticana Species 0.000 description 1
- 235000011330 Armoracia rusticana Nutrition 0.000 description 1
- 108090000101 Asclepain Proteins 0.000 description 1
- DCXYFEDJOCDNAF-UHFFFAOYSA-N Asparagine Natural products OC(=O)C(N)CC(N)=O DCXYFEDJOCDNAF-UHFFFAOYSA-N 0.000 description 1
- 102000004580 Aspartic Acid Proteases Human genes 0.000 description 1
- 102000009422 Aspartic endopeptidases Human genes 0.000 description 1
- 108030004804 Aspartic endopeptidases Proteins 0.000 description 1
- 101710082738 Aspartic protease 3 Proteins 0.000 description 1
- 101000961203 Aspergillus awamori Glucoamylase Proteins 0.000 description 1
- 241000228195 Aspergillus ficuum Species 0.000 description 1
- 241000892910 Aspergillus foetidus Species 0.000 description 1
- 241001480052 Aspergillus japonicus Species 0.000 description 1
- 101900127796 Aspergillus oryzae Glucoamylase Proteins 0.000 description 1
- 101900318521 Aspergillus oryzae Triosephosphate isomerase Proteins 0.000 description 1
- 241000228257 Aspergillus sp. Species 0.000 description 1
- 241001367049 Autographa Species 0.000 description 1
- 101150071434 BAR1 gene Proteins 0.000 description 1
- 101000775727 Bacillus amyloliquefaciens Alpha-amylase Proteins 0.000 description 1
- 241000193752 Bacillus circulans Species 0.000 description 1
- 241001328122 Bacillus clausii Species 0.000 description 1
- 241000193749 Bacillus coagulans Species 0.000 description 1
- 108010029675 Bacillus licheniformis alpha-amylase Proteins 0.000 description 1
- 241000194107 Bacillus megaterium Species 0.000 description 1
- 241000194110 Bacillus sp. (in: Bacteria) Species 0.000 description 1
- 108010045681 Bacillus stearothermophilus neutral protease Proteins 0.000 description 1
- 101900040182 Bacillus subtilis Levansucrase Proteins 0.000 description 1
- 241000193388 Bacillus thuringiensis Species 0.000 description 1
- 108010001478 Bacitracin Proteins 0.000 description 1
- 108010066768 Bacterial leucyl aminopeptidase Proteins 0.000 description 1
- 241000003910 Baronia <angiosperm> Species 0.000 description 1
- 102100030981 Beta-alanine-activating enzyme Human genes 0.000 description 1
- 102100026189 Beta-galactosidase Human genes 0.000 description 1
- 108010015428 Bilirubin oxidase Proteins 0.000 description 1
- LSNNMFCWUKXFEE-UHFFFAOYSA-M Bisulfite Chemical compound OS([O-])=O LSNNMFCWUKXFEE-UHFFFAOYSA-M 0.000 description 1
- 241000235432 Blastocladiella Species 0.000 description 1
- 102000015081 Blood Coagulation Factors Human genes 0.000 description 1
- 108010039209 Blood Coagulation Factors Proteins 0.000 description 1
- BTBUEUYNUDRHOZ-UHFFFAOYSA-N Borate Chemical compound [O-]B([O-])[O-] BTBUEUYNUDRHOZ-UHFFFAOYSA-N 0.000 description 1
- 241000193764 Brevibacillus brevis Species 0.000 description 1
- 101100280051 Brucella abortus biovar 1 (strain 9-941) eryH gene Proteins 0.000 description 1
- 241000235172 Bullera Species 0.000 description 1
- 239000002126 C01EB10 - Adenosine Substances 0.000 description 1
- 102000034342 Calnexin Human genes 0.000 description 1
- 108010056891 Calnexin Proteins 0.000 description 1
- 101100507655 Canis lupus familiaris HSPA1 gene Proteins 0.000 description 1
- 239000005635 Caprylic acid (CAS 124-07-2) Substances 0.000 description 1
- 101710132601 Capsid protein Proteins 0.000 description 1
- BVKZGUZCCUSVTD-UHFFFAOYSA-L Carbonate Chemical compound [O-]C([O-])=O BVKZGUZCCUSVTD-UHFFFAOYSA-L 0.000 description 1
- 102000005367 Carboxypeptidases Human genes 0.000 description 1
- 108010006303 Carboxypeptidases Proteins 0.000 description 1
- 108090000391 Caricain Proteins 0.000 description 1
- 102100035882 Catalase Human genes 0.000 description 1
- 108010053835 Catalase Proteins 0.000 description 1
- 108010031396 Catechol oxidase Proteins 0.000 description 1
- 102000030523 Catechol oxidase Human genes 0.000 description 1
- 241000186221 Cellulosimicrobium cellulans Species 0.000 description 1
- 102100037633 Centrin-3 Human genes 0.000 description 1
- 229920002101 Chitin Polymers 0.000 description 1
- 108010022172 Chitinases Proteins 0.000 description 1
- 102000012286 Chitinases Human genes 0.000 description 1
- ZAMOUSCENKQFHK-UHFFFAOYSA-N Chlorine atom Chemical compound [Cl] ZAMOUSCENKQFHK-UHFFFAOYSA-N 0.000 description 1
- 102000011022 Chorionic Gonadotropin Human genes 0.000 description 1
- 108010062540 Chorionic Gonadotropin Proteins 0.000 description 1
- 241000588881 Chromobacterium Species 0.000 description 1
- 241000146387 Chromobacterium viscosum Species 0.000 description 1
- 108090001069 Chymopapain Proteins 0.000 description 1
- 108090000746 Chymosin Proteins 0.000 description 1
- 108090000317 Chymotrypsin Proteins 0.000 description 1
- 108020004638 Circular DNA Proteins 0.000 description 1
- 101710094648 Coat protein Proteins 0.000 description 1
- 241001279801 Coelomomyces Species 0.000 description 1
- 101800000414 Corticotropin Proteins 0.000 description 1
- 239000000055 Corticotropin-Releasing Hormone Substances 0.000 description 1
- 229920000742 Cotton Polymers 0.000 description 1
- 241000699800 Cricetinae Species 0.000 description 1
- 241000699802 Cricetulus griseus Species 0.000 description 1
- MIKUYHXYGGJMLM-GIMIYPNGSA-N Crotonoside Natural products C1=NC2=C(N)NC(=O)N=C2N1[C@H]1O[C@@H](CO)[C@H](O)[C@@H]1O MIKUYHXYGGJMLM-GIMIYPNGSA-N 0.000 description 1
- 241000490729 Cryptococcaceae Species 0.000 description 1
- 241000221199 Cryptococcus <basidiomycete yeast> Species 0.000 description 1
- 241000047214 Cyclocybe cylindracea Species 0.000 description 1
- 108090000395 Cysteine Endopeptidases Proteins 0.000 description 1
- 102000003950 Cysteine Endopeptidases Human genes 0.000 description 1
- UHDGCWIWMRVCDJ-PSQAKQOGSA-N Cytidine Natural products O=C1N=C(N)C=CN1[C@@H]1[C@@H](O)[C@@H](O)[C@H](CO)O1 UHDGCWIWMRVCDJ-PSQAKQOGSA-N 0.000 description 1
- 102000018832 Cytochromes Human genes 0.000 description 1
- 108010052832 Cytochromes Proteins 0.000 description 1
- 102000004127 Cytokines Human genes 0.000 description 1
- 108090000695 Cytokines Proteins 0.000 description 1
- 102100034560 Cytosol aminopeptidase Human genes 0.000 description 1
- NYHBQMYGNKIUIF-UHFFFAOYSA-N D-guanosine Natural products C1=2NC(N)=NC(=O)C=2N=CN1C1OC(CO)C(O)C1O NYHBQMYGNKIUIF-UHFFFAOYSA-N 0.000 description 1
- 230000004544 DNA amplification Effects 0.000 description 1
- 239000003155 DNA primer Substances 0.000 description 1
- 101710088194 Dehydrogenase Proteins 0.000 description 1
- CKTSBUTUHBMZGZ-UHFFFAOYSA-N Deoxycytidine Natural products O=C1N=C(N)C=CN1C1OC(CO)C(O)C1 CKTSBUTUHBMZGZ-UHFFFAOYSA-N 0.000 description 1
- 206010012438 Dermatitis atopic Diseases 0.000 description 1
- 229920002307 Dextran Polymers 0.000 description 1
- 239000004375 Dextrin Substances 0.000 description 1
- 229920001353 Dextrin Polymers 0.000 description 1
- 101710188975 Dibasic-processing endoprotease Proteins 0.000 description 1
- 101100342470 Dictyostelium discoideum pkbA gene Proteins 0.000 description 1
- 102100024746 Dihydrofolate reductase Human genes 0.000 description 1
- 108090000204 Dipeptidase 1 Proteins 0.000 description 1
- 108700033069 EC 1.97.-.- Proteins 0.000 description 1
- 108700033392 EC 2.1.-.- Proteins 0.000 description 1
- 102000057846 EC 2.1.-.- Human genes 0.000 description 1
- 108700033921 EC 3.4.23.20 Proteins 0.000 description 1
- KCXVZYZYPLLWCC-UHFFFAOYSA-N EDTA Chemical compound OC(=O)CN(CC(O)=O)CCN(CC(O)=O)CC(O)=O KCXVZYZYPLLWCC-UHFFFAOYSA-N 0.000 description 1
- 241000196324 Embryophyta Species 0.000 description 1
- 241000228138 Emericella Species 0.000 description 1
- 108010067770 Endopeptidase K Proteins 0.000 description 1
- 108700041152 Endoplasmic Reticulum Chaperone BiP Proteins 0.000 description 1
- 102100021451 Endoplasmic reticulum chaperone BiP Human genes 0.000 description 1
- BRLQWZUYTZBJKN-UHFFFAOYSA-N Epichlorohydrin Chemical compound ClCC1CO1 BRLQWZUYTZBJKN-UHFFFAOYSA-N 0.000 description 1
- YQYJSBFKSSDGFO-UHFFFAOYSA-N Epihygromycin Natural products OC1C(O)C(C(=O)C)OC1OC(C(=C1)O)=CC=C1C=C(C)C(=O)NC1C(O)C(O)C2OCOC2C1O YQYJSBFKSSDGFO-UHFFFAOYSA-N 0.000 description 1
- 102000003951 Erythropoietin Human genes 0.000 description 1
- 108090000394 Erythropoietin Proteins 0.000 description 1
- 101100385973 Escherichia coli (strain K12) cycA gene Proteins 0.000 description 1
- 241000701959 Escherichia virus Lambda Species 0.000 description 1
- 241001136487 Eurotium Species 0.000 description 1
- 108090000270 Ficain Proteins 0.000 description 1
- 241000221207 Filobasidium Species 0.000 description 1
- 241000192125 Firmicutes Species 0.000 description 1
- 241000145614 Fusarium bactridioides Species 0.000 description 1
- 241000223194 Fusarium culmorum Species 0.000 description 1
- 241000223195 Fusarium graminearum Species 0.000 description 1
- 241001112697 Fusarium reticulatum Species 0.000 description 1
- 241001014439 Fusarium sarcochroum Species 0.000 description 1
- 241000453701 Galactomyces candidum Species 0.000 description 1
- 108010093031 Galactosidases Proteins 0.000 description 1
- 102000002464 Galactosidases Human genes 0.000 description 1
- 108010001515 Galectin 4 Proteins 0.000 description 1
- 102100039556 Galectin-4 Human genes 0.000 description 1
- 101100001650 Geobacillus stearothermophilus amyM gene Proteins 0.000 description 1
- 101100460773 Geobacillus stearothermophilus nprA gene Proteins 0.000 description 1
- 101000930822 Giardia intestinalis Dipeptidyl-peptidase 4 Proteins 0.000 description 1
- 102400000321 Glucagon Human genes 0.000 description 1
- 108060003199 Glucagon Proteins 0.000 description 1
- 108050008938 Glucoamylases Proteins 0.000 description 1
- 239000004366 Glucose oxidase Substances 0.000 description 1
- 108010055629 Glucosyltransferases Proteins 0.000 description 1
- 102000000340 Glucosyltransferases Human genes 0.000 description 1
- 239000005561 Glufosinate Substances 0.000 description 1
- SXRSQZLOMIGNAQ-UHFFFAOYSA-N Glutaraldehyde Chemical compound O=CCCCC=O SXRSQZLOMIGNAQ-UHFFFAOYSA-N 0.000 description 1
- 108010036684 Glycine Dehydrogenase Proteins 0.000 description 1
- 102100033495 Glycine dehydrogenase (decarboxylating), mitochondrial Human genes 0.000 description 1
- 244000068988 Glycine max Species 0.000 description 1
- 235000010469 Glycine max Nutrition 0.000 description 1
- 102000005744 Glycoside Hydrolases Human genes 0.000 description 1
- 108010031186 Glycoside Hydrolases Proteins 0.000 description 1
- 102100021181 Golgi phosphoprotein 3 Human genes 0.000 description 1
- 108010051696 Growth Hormone Proteins 0.000 description 1
- 101150009006 HIS3 gene Proteins 0.000 description 1
- 101150112743 HSPA5 gene Proteins 0.000 description 1
- 101100295959 Halobacterium salinarum (strain ATCC 700922 / JCM 11081 / NRC-1) arcB gene Proteins 0.000 description 1
- 101100246753 Halobacterium salinarum (strain ATCC 700922 / JCM 11081 / NRC-1) pyrF gene Proteins 0.000 description 1
- 240000001194 Heliotropium europaeum Species 0.000 description 1
- 229920002488 Hemicellulose Polymers 0.000 description 1
- SQUHHTBVTRBESD-UHFFFAOYSA-N Hexa-Ac-myo-Inositol Natural products CC(=O)OC1C(OC(C)=O)C(OC(C)=O)C(OC(C)=O)C(OC(C)=O)C1OC(C)=O SQUHHTBVTRBESD-UHFFFAOYSA-N 0.000 description 1
- 241000282412 Homo Species 0.000 description 1
- 101000773364 Homo sapiens Beta-alanine-activating enzyme Proteins 0.000 description 1
- 101000880522 Homo sapiens Centrin-3 Proteins 0.000 description 1
- 241000291718 Hoplocampa brevis Species 0.000 description 1
- 108010001336 Horseradish Peroxidase Proteins 0.000 description 1
- 108090000144 Human Proteins Proteins 0.000 description 1
- 102000003839 Human Proteins Human genes 0.000 description 1
- 241000701109 Human adenovirus 2 Species 0.000 description 1
- 241001135569 Human adenovirus 5 Species 0.000 description 1
- 102000004157 Hydrolases Human genes 0.000 description 1
- 108090000604 Hydrolases Proteins 0.000 description 1
- 108700002232 Immediate-Early Genes Proteins 0.000 description 1
- 102000001706 Immunoglobulin Fab Fragments Human genes 0.000 description 1
- 108010054477 Immunoglobulin Fab Fragments Proteins 0.000 description 1
- IMQLKJBTEOYOSI-GPIVLXJGSA-N Inositol-hexakisphosphate Chemical compound OP(O)(=O)O[C@H]1[C@H](OP(O)(O)=O)[C@@H](OP(O)(O)=O)[C@H](OP(O)(O)=O)[C@H](OP(O)(O)=O)[C@@H]1OP(O)(O)=O IMQLKJBTEOYOSI-GPIVLXJGSA-N 0.000 description 1
- 108700001097 Insect Genes Proteins 0.000 description 1
- 108090001061 Insulin Proteins 0.000 description 1
- 102000004877 Insulin Human genes 0.000 description 1
- 108090000723 Insulin-Like Growth Factor I Proteins 0.000 description 1
- 108010050904 Interferons Proteins 0.000 description 1
- 102000014150 Interferons Human genes 0.000 description 1
- 239000007836 KH2PO4 Substances 0.000 description 1
- 241001138401 Kluyveromyces lactis Species 0.000 description 1
- 241000235058 Komagataella pastoris Species 0.000 description 1
- 108010008292 L-Amino Acid Oxidase Proteins 0.000 description 1
- ONIBWKKTOPOVIA-BYPYZUCNSA-N L-Proline Chemical compound OC(=O)[C@@H]1CCCN1 ONIBWKKTOPOVIA-BYPYZUCNSA-N 0.000 description 1
- QNAYBMKLOCPYGJ-REOHCLBHSA-N L-alanine Chemical compound C[C@H](N)C(O)=O QNAYBMKLOCPYGJ-REOHCLBHSA-N 0.000 description 1
- 108030000198 L-amino-acid dehydrogenases Proteins 0.000 description 1
- 102000007070 L-amino-acid oxidase Human genes 0.000 description 1
- DCXYFEDJOCDNAF-REOHCLBHSA-N L-asparagine Chemical compound OC(=O)[C@@H](N)CC(N)=O DCXYFEDJOCDNAF-REOHCLBHSA-N 0.000 description 1
- 108030000910 L-aspartate oxidases Proteins 0.000 description 1
- CKLJMWTZIZZHCS-REOHCLBHSA-N L-aspartic acid Chemical compound OC(=O)[C@@H](N)CC(O)=O CKLJMWTZIZZHCS-REOHCLBHSA-N 0.000 description 1
- 108010069325 L-glutamate oxidase Proteins 0.000 description 1
- WHUUTDBJXJRKMK-VKHMYHEASA-N L-glutamic acid Chemical compound OC(=O)[C@@H](N)CCC(O)=O WHUUTDBJXJRKMK-VKHMYHEASA-N 0.000 description 1
- 108010004733 L-lysine oxidase Proteins 0.000 description 1
- AYFVYJQAPQTCCC-GBXIJSLDSA-N L-threonine Chemical compound C[C@@H](O)[C@H](N)C(O)=O AYFVYJQAPQTCCC-GBXIJSLDSA-N 0.000 description 1
- 241000235087 Lachancea kluyveri Species 0.000 description 1
- 108010059881 Lactase Proteins 0.000 description 1
- 108091026898 Leader sequence (mRNA) Proteins 0.000 description 1
- 241000255777 Lepidoptera Species 0.000 description 1
- 241000321520 Leptomitales Species 0.000 description 1
- 108010028658 Leucine Dehydrogenase Proteins 0.000 description 1
- 101710098556 Lipase A Proteins 0.000 description 1
- 101710098554 Lipase B Proteins 0.000 description 1
- 102000003820 Lipoxygenases Human genes 0.000 description 1
- 108090000128 Lipoxygenases Proteins 0.000 description 1
- 108010048733 Lipozyme Proteins 0.000 description 1
- 102000009151 Luteinizing Hormone Human genes 0.000 description 1
- 108010073521 Luteinizing Hormone Proteins 0.000 description 1
- 101710099648 Lysosomal acid lipase/cholesteryl ester hydrolase Proteins 0.000 description 1
- 102100026001 Lysosomal acid lipase/cholesteryl ester hydrolase Human genes 0.000 description 1
- 101150068888 MET3 gene Proteins 0.000 description 1
- 101710125418 Major capsid protein Proteins 0.000 description 1
- GUBGYTABKSRVRQ-PICCSMPSSA-N Maltose Natural products O[C@@H]1[C@@H](O)[C@H](O)[C@@H](CO)O[C@@H]1O[C@@H]1[C@@H](CO)OC(O)[C@H](O)[C@H]1O GUBGYTABKSRVRQ-PICCSMPSSA-N 0.000 description 1
- 229920000057 Mannan Polymers 0.000 description 1
- 108090000131 Metalloendopeptidases Proteins 0.000 description 1
- 102000003843 Metalloendopeptidases Human genes 0.000 description 1
- 108090000192 Methionyl aminopeptidases Proteins 0.000 description 1
- 102000034452 Methionyl aminopeptidases Human genes 0.000 description 1
- 108010014251 Muramidase Proteins 0.000 description 1
- 102000016943 Muramidase Human genes 0.000 description 1
- 241000699666 Mus <mouse, genus> Species 0.000 description 1
- 101100379123 Mus musculus Npr1 gene Proteins 0.000 description 1
- 101100235161 Mycolicibacterium smegmatis (strain ATCC 700084 / mc(2)155) lerI gene Proteins 0.000 description 1
- 241001674208 Mycothermus thermophilus Species 0.000 description 1
- 108010062010 N-Acetylmuramoyl-L-alanine Amidase Proteins 0.000 description 1
- 229930193140 Neomycin Natural products 0.000 description 1
- 241000221961 Neurospora crassa Species 0.000 description 1
- 101100022915 Neurospora crassa (strain ATCC 24698 / 74-OR23-1A / CBS 708.71 / DSM 1257 / FGSC 987) cys-11 gene Proteins 0.000 description 1
- 108090000913 Nitrate Reductases Proteins 0.000 description 1
- 101710141454 Nucleoprotein Proteins 0.000 description 1
- BXEFQPCKQSTMKA-UHFFFAOYSA-N OC(=O)C=[N+]=[N-] Chemical compound OC(=O)C=[N+]=[N-] BXEFQPCKQSTMKA-UHFFFAOYSA-N 0.000 description 1
- 108020005187 Oligonucleotide Probes Proteins 0.000 description 1
- 102000007981 Ornithine carbamoyltransferase Human genes 0.000 description 1
- 101710113020 Ornithine transcarbamylase, mitochondrial Proteins 0.000 description 1
- 102100037214 Orotidine 5'-phosphate decarboxylase Human genes 0.000 description 1
- 108010055012 Orotidine-5'-phosphate decarboxylase Proteins 0.000 description 1
- 101100378536 Ovis aries ADRB1 gene Proteins 0.000 description 1
- 229910019142 PO4 Inorganic materials 0.000 description 1
- 241001236817 Paecilomyces <Clavicipitaceae> Species 0.000 description 1
- 241000194109 Paenibacillus lautus Species 0.000 description 1
- 108090000526 Papain Proteins 0.000 description 1
- 102000003982 Parathyroid hormone Human genes 0.000 description 1
- 108090000445 Parathyroid hormone Proteins 0.000 description 1
- 206010034133 Pathogen resistance Diseases 0.000 description 1
- 241001494479 Pecora Species 0.000 description 1
- 108010029182 Pectin lyase Proteins 0.000 description 1
- 244000271379 Penicillium camembertii Species 0.000 description 1
- 108010067902 Peptide Library Proteins 0.000 description 1
- 108700020962 Peroxidase Proteins 0.000 description 1
- YGYAWVDWMABLBF-UHFFFAOYSA-N Phosgene Chemical compound ClC(Cl)=O YGYAWVDWMABLBF-UHFFFAOYSA-N 0.000 description 1
- 108090000608 Phosphoric Monoester Hydrolases Proteins 0.000 description 1
- 102000004160 Phosphoric Monoester Hydrolases Human genes 0.000 description 1
- OAICVXFJPJFONN-UHFFFAOYSA-N Phosphorus Chemical compound [P] OAICVXFJPJFONN-UHFFFAOYSA-N 0.000 description 1
- 241000425347 Phyla <beetle> Species 0.000 description 1
- 241000224486 Physarum polycephalum Species 0.000 description 1
- 108010001014 Plasminogen Activators Proteins 0.000 description 1
- 102000001938 Plasminogen Activators Human genes 0.000 description 1
- 241000276498 Pollachius virens Species 0.000 description 1
- 239000004698 Polyethylene Substances 0.000 description 1
- 101710182846 Polyhedrin Proteins 0.000 description 1
- 241000789035 Polyporus pinsitus Species 0.000 description 1
- 102100027467 Pro-opiomelanocortin Human genes 0.000 description 1
- 101710083689 Probable capsid protein Proteins 0.000 description 1
- 101710093543 Probable non-specific lipid-transfer protein Proteins 0.000 description 1
- 102000003946 Prolactin Human genes 0.000 description 1
- 108010057464 Prolactin Proteins 0.000 description 1
- ONIBWKKTOPOVIA-UHFFFAOYSA-N Proline Natural products OC(=O)C1CCCN1 ONIBWKKTOPOVIA-UHFFFAOYSA-N 0.000 description 1
- 229940124158 Protease/peptidase inhibitor Drugs 0.000 description 1
- 102000003866 Protein Disulfide Reductase (Glutathione) Human genes 0.000 description 1
- 108090000213 Protein Disulfide Reductase (Glutathione) Proteins 0.000 description 1
- 108050004742 Protein disulphide isomerases Proteins 0.000 description 1
- 102000016227 Protein disulphide isomerases Human genes 0.000 description 1
- 108010003894 Protein-Lysine 6-Oxidase Proteins 0.000 description 1
- 102000004669 Protein-Lysine 6-Oxidase Human genes 0.000 description 1
- 102100038094 Protein-glutamine gamma-glutamyltransferase E Human genes 0.000 description 1
- 101710182788 Protein-glutamine gamma-glutamyltransferase E Proteins 0.000 description 1
- 241000589774 Pseudomonas sp. Species 0.000 description 1
- 241000221535 Pucciniales Species 0.000 description 1
- 241001465752 Purpureocillium lilacinum Species 0.000 description 1
- 101710185622 Pyrrolidone-carboxylate peptidase Proteins 0.000 description 1
- 230000004570 RNA-binding Effects 0.000 description 1
- 241000700159 Rattus Species 0.000 description 1
- 108090000103 Relaxin Proteins 0.000 description 1
- 102000003743 Relaxin Human genes 0.000 description 1
- 241000303962 Rhizopus delemar Species 0.000 description 1
- 240000005384 Rhizopus oryzae Species 0.000 description 1
- 101100394989 Rhodopseudomonas palustris (strain ATCC BAA-98 / CGA009) hisI gene Proteins 0.000 description 1
- 241000223252 Rhodotorula Species 0.000 description 1
- 108091028664 Ribonucleotide Proteins 0.000 description 1
- 244000157378 Rubus niveus Species 0.000 description 1
- 101000718529 Saccharolobus solfataricus (strain ATCC 35092 / DSM 1617 / JCM 11322 / P2) Alpha-galactosidase Proteins 0.000 description 1
- 235000003534 Saccharomyces carlsbergensis Nutrition 0.000 description 1
- 101100111629 Saccharomyces cerevisiae (strain ATCC 204508 / S288c) KAR2 gene Proteins 0.000 description 1
- 101900354623 Saccharomyces cerevisiae Galactokinase Proteins 0.000 description 1
- 101001035673 Saccharomyces cerevisiae Heme-responsive zinc finger transcription factor HAP1 Proteins 0.000 description 1
- 235000001006 Saccharomyces cerevisiae var diastaticus Nutrition 0.000 description 1
- 244000206963 Saccharomyces cerevisiae var. diastaticus Species 0.000 description 1
- 241000204893 Saccharomyces douglasii Species 0.000 description 1
- 241001407717 Saccharomyces norbensis Species 0.000 description 1
- 241001123227 Saccharomyces pastorianus Species 0.000 description 1
- 241000235344 Saccharomycetaceae Species 0.000 description 1
- 241000235343 Saccharomycetales Species 0.000 description 1
- 241001326564 Saccharomycotina Species 0.000 description 1
- 108090000077 Saccharopepsin Proteins 0.000 description 1
- 241000235347 Schizosaccharomyces pombe Species 0.000 description 1
- 101100097319 Schizosaccharomyces pombe (strain 972 / ATCC 24843) ala1 gene Proteins 0.000 description 1
- 101100022918 Schizosaccharomyces pombe (strain 972 / ATCC 24843) sua1 gene Proteins 0.000 description 1
- 102000003667 Serine Endopeptidases Human genes 0.000 description 1
- 108090000083 Serine Endopeptidases Proteins 0.000 description 1
- 102000013275 Somatomedins Human genes 0.000 description 1
- 108010056088 Somatostatin Proteins 0.000 description 1
- 102000005157 Somatostatin Human genes 0.000 description 1
- 102100038803 Somatotropin Human genes 0.000 description 1
- 241000256251 Spodoptera frugiperda Species 0.000 description 1
- 241000228389 Sporidiobolus Species 0.000 description 1
- 229920002472 Starch Polymers 0.000 description 1
- 108010090804 Streptavidin Proteins 0.000 description 1
- 101100309436 Streptococcus mutans serotype c (strain ATCC 700610 / UA159) ftf gene Proteins 0.000 description 1
- 108010023197 Streptokinase Proteins 0.000 description 1
- 241000520730 Streptomyces cinnamoneus Species 0.000 description 1
- 241000187432 Streptomyces coelicolor Species 0.000 description 1
- 101100370749 Streptomyces coelicolor (strain ATCC BAA-471 / A3(2) / M145) trpC1 gene Proteins 0.000 description 1
- 241000499056 Streptomyces griseocarneus Species 0.000 description 1
- 241000187391 Streptomyces hygroscopicus Species 0.000 description 1
- 241000187389 Streptomyces lavendulae Species 0.000 description 1
- 241000969738 Streptomyces libani Species 0.000 description 1
- 241000187398 Streptomyces lividans Species 0.000 description 1
- 241000218483 Streptomyces lydicus Species 0.000 description 1
- 241001495137 Streptomyces mobaraensis Species 0.000 description 1
- 241001468239 Streptomyces murinus Species 0.000 description 1
- 241000187180 Streptomyces sp. Species 0.000 description 1
- 101710173714 Subtilisin amylosacchariticus Proteins 0.000 description 1
- 229930006000 Sucrose Natural products 0.000 description 1
- CZMRCDWAGMRECN-UGDNZRGBSA-N Sucrose Chemical compound O[C@H]1[C@H](O)[C@@H](CO)O[C@@]1(CO)O[C@@H]1[C@H](O)[C@@H](O)[C@H](O)[C@@H](CO)O1 CZMRCDWAGMRECN-UGDNZRGBSA-N 0.000 description 1
- QAOWNCQODCNURD-UHFFFAOYSA-L Sulfate Chemical compound [O-]S([O-])(=O)=O QAOWNCQODCNURD-UHFFFAOYSA-L 0.000 description 1
- QAOWNCQODCNURD-UHFFFAOYSA-N Sulfuric acid Chemical compound OS(O)(=O)=O QAOWNCQODCNURD-UHFFFAOYSA-N 0.000 description 1
- 102000019197 Superoxide Dismutase Human genes 0.000 description 1
- 108010012715 Superoxide dismutase Proteins 0.000 description 1
- 230000006044 T cell activation Effects 0.000 description 1
- 108091008874 T cell receptors Proteins 0.000 description 1
- 102000016266 T-Cell Antigen Receptors Human genes 0.000 description 1
- 241001540751 Talaromyces ruber Species 0.000 description 1
- 239000004098 Tetracycline Substances 0.000 description 1
- 101100157012 Thermoanaerobacterium saccharolyticum (strain DSM 8691 / JW/SL-YS485) xynB gene Proteins 0.000 description 1
- 102000013090 Thioredoxin-Disulfide Reductase Human genes 0.000 description 1
- 108010079911 Thioredoxin-disulfide reductase Proteins 0.000 description 1
- 108091036066 Three prime untranslated region Proteins 0.000 description 1
- AYFVYJQAPQTCCC-UHFFFAOYSA-N Threonine Natural products CC(O)C(N)C(O)=O AYFVYJQAPQTCCC-UHFFFAOYSA-N 0.000 description 1
- 239000004473 Threonine Substances 0.000 description 1
- 108010046075 Thymosin Proteins 0.000 description 1
- 102000007501 Thymosin Human genes 0.000 description 1
- 108010061174 Thyrotropin Proteins 0.000 description 1
- 102000011923 Thyrotropin Human genes 0.000 description 1
- ISWQCIVKKSOKNN-UHFFFAOYSA-L Tiron Chemical compound [Na+].[Na+].OC1=CC(S([O-])(=O)=O)=CC(S([O-])(=O)=O)=C1O ISWQCIVKKSOKNN-UHFFFAOYSA-L 0.000 description 1
- 108090000373 Tissue Plasminogen Activator Proteins 0.000 description 1
- 102000003978 Tissue Plasminogen Activator Human genes 0.000 description 1
- 108091023040 Transcription factor Proteins 0.000 description 1
- 241000378866 Trichoderma koningii Species 0.000 description 1
- 241000223262 Trichoderma longibrachiatum Species 0.000 description 1
- 241000223261 Trichoderma viride Species 0.000 description 1
- 241000255993 Trichoplusia ni Species 0.000 description 1
- 102000005924 Triose-Phosphate Isomerase Human genes 0.000 description 1
- 108700015934 Triose-phosphate isomerases Proteins 0.000 description 1
- 239000007983 Tris buffer Substances 0.000 description 1
- 108030000963 Tryptophanyl aminopeptidases Proteins 0.000 description 1
- 101150050575 URA3 gene Proteins 0.000 description 1
- 108010009135 Uca pugilator serine collagenase 1 Proteins 0.000 description 1
- 208000024780 Urticaria Diseases 0.000 description 1
- 241000221561 Ustilaginales Species 0.000 description 1
- 101710100604 Valine dehydrogenase Proteins 0.000 description 1
- 108010004977 Vasopressins Proteins 0.000 description 1
- 102000002852 Vasopressins Human genes 0.000 description 1
- 108020005202 Viral DNA Proteins 0.000 description 1
- 108010038900 X-Pro aminopeptidase Proteins 0.000 description 1
- 238000010521 absorption reaction Methods 0.000 description 1
- 238000003916 acid precipitation Methods 0.000 description 1
- 150000007513 acids Chemical class 0.000 description 1
- 108090000350 actinidain Proteins 0.000 description 1
- 230000010933 acylation Effects 0.000 description 1
- 238000005917 acylation reaction Methods 0.000 description 1
- 108700014220 acyltransferase activity proteins Proteins 0.000 description 1
- 229960005305 adenosine Drugs 0.000 description 1
- 238000001042 affinity chromatography Methods 0.000 description 1
- 238000001261 affinity purification Methods 0.000 description 1
- 239000008272 agar Substances 0.000 description 1
- 108010045649 agarase Proteins 0.000 description 1
- 239000011543 agarose gel Substances 0.000 description 1
- 238000007818 agglutination assay Methods 0.000 description 1
- 230000002776 aggregation Effects 0.000 description 1
- 238000004220 aggregation Methods 0.000 description 1
- 235000004279 alanine Nutrition 0.000 description 1
- 150000001298 alcohols Chemical class 0.000 description 1
- HAXFWIACAGNFHA-UHFFFAOYSA-N aldrithiol Chemical compound C=1C=CC=NC=1SSC1=CC=CC=N1 HAXFWIACAGNFHA-UHFFFAOYSA-N 0.000 description 1
- 125000000217 alkyl group Chemical group 0.000 description 1
- 229940045714 alkyl sulfonate alkylating agent Drugs 0.000 description 1
- 150000008052 alkyl sulfonates Chemical class 0.000 description 1
- NNISLDGFPWIBDF-MPRBLYSKSA-N alpha-D-Gal-(1->3)-beta-D-Gal-(1->4)-D-GlcNAc Chemical compound O[C@@H]1[C@@H](NC(=O)C)C(O)O[C@H](CO)[C@H]1O[C@H]1[C@H](O)[C@@H](O[C@@H]2[C@@H]([C@@H](O)[C@@H](O)[C@@H](CO)O2)O)[C@@H](O)[C@@H](CO)O1 NNISLDGFPWIBDF-MPRBLYSKSA-N 0.000 description 1
- 150000001412 amines Chemical class 0.000 description 1
- 229910021529 ammonia Inorganic materials 0.000 description 1
- BFNBIHQBYMNNAN-UHFFFAOYSA-N ammonium sulfate Chemical compound N.N.OS(O)(=O)=O BFNBIHQBYMNNAN-UHFFFAOYSA-N 0.000 description 1
- 229910052921 ammonium sulfate Inorganic materials 0.000 description 1
- 238000012870 ammonium sulfate precipitation Methods 0.000 description 1
- 239000001166 ammonium sulphate Substances 0.000 description 1
- 235000011130 ammonium sulphate Nutrition 0.000 description 1
- 229960000723 ampicillin Drugs 0.000 description 1
- AVKUERGKIZMTKX-NJBDSQKTSA-N ampicillin Chemical compound C1([C@@H](N)C(=O)N[C@H]2[C@H]3SC([C@@H](N3C2=O)C(O)=O)(C)C)=CC=CC=C1 AVKUERGKIZMTKX-NJBDSQKTSA-N 0.000 description 1
- 208000003455 anaphylaxis Diseases 0.000 description 1
- 210000004102 animal cell Anatomy 0.000 description 1
- 230000003042 antagnostic effect Effects 0.000 description 1
- 210000000628 antibody-producing cell Anatomy 0.000 description 1
- 101150009206 aprE gene Proteins 0.000 description 1
- PYMYPHUHKUWMLA-UHFFFAOYSA-N arabinose Natural products OCC(O)C(O)C(O)C=O PYMYPHUHKUWMLA-UHFFFAOYSA-N 0.000 description 1
- 101150008194 argB gene Proteins 0.000 description 1
- ODKSFYDXXFIFQN-UHFFFAOYSA-N arginine Natural products OC(=O)C(N)CCCNC(N)=N ODKSFYDXXFIFQN-UHFFFAOYSA-N 0.000 description 1
- 150000004982 aromatic amines Chemical class 0.000 description 1
- 210000004507 artificial chromosome Anatomy 0.000 description 1
- 125000003118 aryl group Chemical group 0.000 description 1
- 125000005228 aryl sulfonate group Chemical group 0.000 description 1
- 235000009582 asparagine Nutrition 0.000 description 1
- 229960001230 asparagine Drugs 0.000 description 1
- 235000003704 aspartic acid Nutrition 0.000 description 1
- 108090000987 aspergillopepsin I Proteins 0.000 description 1
- 238000002820 assay format Methods 0.000 description 1
- 208000006673 asthma Diseases 0.000 description 1
- 125000004429 atom Chemical group 0.000 description 1
- 201000008937 atopic dermatitis Diseases 0.000 description 1
- 229940054340 bacillus coagulans Drugs 0.000 description 1
- 229940097012 bacillus thuringiensis Drugs 0.000 description 1
- 229960003071 bacitracin Drugs 0.000 description 1
- 229930184125 bacitracin Natural products 0.000 description 1
- CLKOFPXJLQSYAH-ABRJDSQDSA-N bacitracin A Chemical compound C1SC([C@@H](N)[C@@H](C)CC)=N[C@@H]1C(=O)N[C@@H](CC(C)C)C(=O)N[C@H](CCC(O)=O)C(=O)N[C@@H]([C@@H](C)CC)C(=O)N[C@@H]1C(=O)N[C@H](CCCN)C(=O)N[C@@H]([C@@H](C)CC)C(=O)N[C@H](CC=2C=CC=CC=2)C(=O)N[C@@H](CC=2N=CNC=2)C(=O)N[C@H](CC(O)=O)C(=O)N[C@@H](CC(N)=O)C(=O)NCCCC1 CLKOFPXJLQSYAH-ABRJDSQDSA-N 0.000 description 1
- 230000009286 beneficial effect Effects 0.000 description 1
- SRBFZHDQGSBBOR-UHFFFAOYSA-N beta-D-Pyranose-Lyxose Natural products OC1COC(O)C(O)C1O SRBFZHDQGSBBOR-UHFFFAOYSA-N 0.000 description 1
- 108010051210 beta-Fructofuranosidase Proteins 0.000 description 1
- 108010005774 beta-Galactosidase Proteins 0.000 description 1
- 108010047754 beta-Glucosidase Proteins 0.000 description 1
- 102000006995 beta-Glucosidase Human genes 0.000 description 1
- DRTQHJPVMGBUCF-PSQAKQOGSA-N beta-L-uridine Natural products O[C@H]1[C@@H](O)[C@H](CO)O[C@@H]1N1C(=O)NC(=O)C=C1 DRTQHJPVMGBUCF-PSQAKQOGSA-N 0.000 description 1
- OQFSQFPPLPISGP-UHFFFAOYSA-N beta-carboxyaspartic acid Natural products OC(=O)C(N)C(C(O)=O)C(O)=O OQFSQFPPLPISGP-UHFFFAOYSA-N 0.000 description 1
- 239000003139 biocide Substances 0.000 description 1
- 210000004369 blood Anatomy 0.000 description 1
- 239000008280 blood Substances 0.000 description 1
- 239000003114 blood coagulation factor Substances 0.000 description 1
- 229940019700 blood coagulation factors Drugs 0.000 description 1
- 210000004899 c-terminal region Anatomy 0.000 description 1
- 239000001506 calcium phosphate Substances 0.000 description 1
- 229910000389 calcium phosphate Inorganic materials 0.000 description 1
- 235000011010 calcium phosphates Nutrition 0.000 description 1
- 244000309466 calf Species 0.000 description 1
- 102000028861 calmodulin binding Human genes 0.000 description 1
- 108091000084 calmodulin binding Proteins 0.000 description 1
- 235000013877 carbamide Nutrition 0.000 description 1
- 235000014633 carbohydrates Nutrition 0.000 description 1
- 125000002915 carbonyl group Chemical group [*:2]C([*:1])=O 0.000 description 1
- 125000002843 carboxylic acid group Chemical group 0.000 description 1
- 230000003197 catalytic effect Effects 0.000 description 1
- 230000034303 cell budding Effects 0.000 description 1
- 239000013592 cell lysate Substances 0.000 description 1
- 230000036755 cellular response Effects 0.000 description 1
- 239000002738 chelating agent Substances 0.000 description 1
- 238000012412 chemical coupling Methods 0.000 description 1
- 229960005091 chloramphenicol Drugs 0.000 description 1
- WIIZWVCIJKGZOK-RKDXNWHRSA-N chloramphenicol Chemical compound ClC(Cl)C(=O)N[C@H](CO)[C@H](O)C1=CC=C([N+]([O-])=O)C=C1 WIIZWVCIJKGZOK-RKDXNWHRSA-N 0.000 description 1
- 239000000460 chlorine Substances 0.000 description 1
- 229910052801 chlorine Inorganic materials 0.000 description 1
- 229940015047 chorionic gonadotropin Drugs 0.000 description 1
- 238000011098 chromatofocusing Methods 0.000 description 1
- 239000013611 chromosomal DNA Substances 0.000 description 1
- 229960002976 chymopapain Drugs 0.000 description 1
- 229960002376 chymotrypsin Drugs 0.000 description 1
- 238000003776 cleavage reaction Methods 0.000 description 1
- 230000000295 complement effect Effects 0.000 description 1
- 238000010276 construction Methods 0.000 description 1
- 238000007796 conventional method Methods 0.000 description 1
- IDLFZVILOHSSID-OVLDLUHVSA-N corticotropin Chemical compound C([C@@H](C(=O)N[C@@H](CO)C(=O)N[C@@H](CCSC)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CC=1NC=NC=1)C(=O)N[C@@H](CC=1C=CC=CC=1)C(=O)N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CC=1C2=CC=CC=C2NC=1)C(=O)NCC(=O)N[C@@H](CCCCN)C(=O)N1[C@@H](CCC1)C(=O)N[C@@H](C(C)C)C(=O)NCC(=O)N[C@@H](CCCCN)C(=O)N[C@@H](CCCCN)C(=O)N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CCCNC(N)=N)C(=O)N1[C@@H](CCC1)C(=O)N[C@@H](C(C)C)C(=O)N[C@@H](CCCCN)C(=O)N[C@@H](C(C)C)C(=O)N[C@@H](CC=1C=CC(O)=CC=1)C(=O)N1[C@@H](CCC1)C(=O)N[C@@H](CC(N)=O)C(=O)NCC(=O)N[C@@H](C)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CC(O)=O)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CO)C(=O)N[C@@H](C)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](C)C(=O)N[C@@H](CC=1C=CC=CC=1)C(=O)N1[C@@H](CCC1)C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CC=1C=CC=CC=1)C(O)=O)NC(=O)[C@@H](N)CO)C1=CC=C(O)C=C1 IDLFZVILOHSSID-OVLDLUHVSA-N 0.000 description 1
- 229960000258 corticotropin Drugs 0.000 description 1
- 239000006071 cream Substances 0.000 description 1
- 238000002447 crystallographic data Methods 0.000 description 1
- 108090000200 cucumisin Proteins 0.000 description 1
- 108010005400 cutinase Proteins 0.000 description 1
- ATDGTVJJHBUTRL-UHFFFAOYSA-N cyanogen bromide Chemical compound BrC#N ATDGTVJJHBUTRL-UHFFFAOYSA-N 0.000 description 1
- 125000004122 cyclic group Chemical group 0.000 description 1
- 235000018417 cysteine Nutrition 0.000 description 1
- XUJNEKJLAYXESH-UHFFFAOYSA-N cysteine Natural products SCC(N)C(O)=O XUJNEKJLAYXESH-UHFFFAOYSA-N 0.000 description 1
- UHDGCWIWMRVCDJ-ZAKLUEHWSA-N cytidine Chemical compound O=C1N=C(N)C=CN1[C@H]1[C@H](O)[C@@H](O)[C@H](CO)O1 UHDGCWIWMRVCDJ-ZAKLUEHWSA-N 0.000 description 1
- 230000009089 cytolysis Effects 0.000 description 1
- 101150005799 dagA gene Proteins 0.000 description 1
- 239000007857 degradation product Substances 0.000 description 1
- 230000003413 degradative effect Effects 0.000 description 1
- 238000004925 denaturation Methods 0.000 description 1
- 230000036425 denaturation Effects 0.000 description 1
- 239000005549 deoxyribonucleoside Substances 0.000 description 1
- 239000005547 deoxyribonucleotide Substances 0.000 description 1
- 125000002637 deoxyribonucleotide group Chemical group 0.000 description 1
- 230000001066 destructive effect Effects 0.000 description 1
- 238000011161 development Methods 0.000 description 1
- 230000018109 developmental process Effects 0.000 description 1
- 235000019425 dextrin Nutrition 0.000 description 1
- 238000000502 dialysis Methods 0.000 description 1
- RAABOESOVLLHRU-UHFFFAOYSA-N diazene Chemical compound N=N RAABOESOVLLHRU-UHFFFAOYSA-N 0.000 description 1
- 229910000071 diazene Inorganic materials 0.000 description 1
- 239000012954 diazonium Substances 0.000 description 1
- 150000001989 diazonium salts Chemical class 0.000 description 1
- 230000004069 differentiation Effects 0.000 description 1
- 238000009792 diffusion process Methods 0.000 description 1
- 239000013024 dilution buffer Substances 0.000 description 1
- ZPWVASYFFYYZEW-UHFFFAOYSA-L dipotassium hydrogen phosphate Chemical compound [K+].[K+].OP([O-])([O-])=O ZPWVASYFFYYZEW-UHFFFAOYSA-L 0.000 description 1
- 229910000396 dipotassium phosphate Inorganic materials 0.000 description 1
- 235000019797 dipotassium phosphate Nutrition 0.000 description 1
- 230000008034 disappearance Effects 0.000 description 1
- 201000010099 disease Diseases 0.000 description 1
- 208000037265 diseases, disorders, signs and symptoms Diseases 0.000 description 1
- 238000004851 dishwashing Methods 0.000 description 1
- 238000005315 distribution function Methods 0.000 description 1
- 231100000673 dose–response relationship Toxicity 0.000 description 1
- 230000009977 dual effect Effects 0.000 description 1
- 238000004520 electroporation Methods 0.000 description 1
- 210000002472 endoplasmic reticulum Anatomy 0.000 description 1
- 238000001952 enzyme assay Methods 0.000 description 1
- 229940105423 erythropoietin Drugs 0.000 description 1
- 125000004185 ester group Chemical group 0.000 description 1
- 230000005284 excitation Effects 0.000 description 1
- 230000007717 exclusion Effects 0.000 description 1
- 230000001747 exhibiting effect Effects 0.000 description 1
- 108010093305 exopolygalacturonase Proteins 0.000 description 1
- 238000000605 extraction Methods 0.000 description 1
- 235000019836 ficin Nutrition 0.000 description 1
- 239000000706 filtrate Substances 0.000 description 1
- 239000012530 fluid Substances 0.000 description 1
- 235000013305 food Nutrition 0.000 description 1
- 230000037406 food intake Effects 0.000 description 1
- 238000009472 formulation Methods 0.000 description 1
- 229930182830 galactose Natural products 0.000 description 1
- 238000001641 gel filtration chromatography Methods 0.000 description 1
- MASNOZXLGMXCHN-ZLPAWPGGSA-N glucagon Chemical compound C([C@@H](C(=O)N[C@H](C(=O)N[C@@H](CCC(N)=O)C(=O)N[C@@H](CC=1C2=CC=CC=C2NC=1)C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](CCSC)C(=O)N[C@@H](CC(N)=O)C(=O)N[C@@H]([C@@H](C)O)C(O)=O)C(C)C)NC(=O)[C@H](CC(O)=O)NC(=O)[C@H](CCC(N)=O)NC(=O)[C@H](C)NC(=O)[C@H](CCCNC(N)=N)NC(=O)[C@H](CCCNC(N)=N)NC(=O)[C@H](CO)NC(=O)[C@H](CC(O)=O)NC(=O)[C@H](CC(C)C)NC(=O)[C@H](CC=1C=CC(O)=CC=1)NC(=O)[C@H](CCCCN)NC(=O)[C@H](CO)NC(=O)[C@H](CC=1C=CC(O)=CC=1)NC(=O)[C@H](CC(O)=O)NC(=O)[C@H](CO)NC(=O)[C@@H](NC(=O)[C@H](CC=1C=CC=CC=1)NC(=O)[C@@H](NC(=O)CNC(=O)[C@H](CCC(N)=O)NC(=O)[C@H](CO)NC(=O)[C@@H](N)CC=1NC=NC=1)[C@@H](C)O)[C@@H](C)O)C1=CC=CC=C1 MASNOZXLGMXCHN-ZLPAWPGGSA-N 0.000 description 1
- 229960004666 glucagon Drugs 0.000 description 1
- 229940116332 glucose oxidase Drugs 0.000 description 1
- 125000000291 glutamic acid group Chemical group N[C@@H](CCC(O)=O)C(=O)* 0.000 description 1
- 125000000404 glutamine group Chemical group N[C@@H](CCC(N)=O)C(=O)* 0.000 description 1
- 102000006602 glyceraldehyde-3-phosphate dehydrogenase Human genes 0.000 description 1
- 230000002414 glycolytic effect Effects 0.000 description 1
- 102000035122 glycosylated proteins Human genes 0.000 description 1
- 108091005608 glycosylated proteins Proteins 0.000 description 1
- 230000013595 glycosylation Effects 0.000 description 1
- 238000006206 glycosylation reaction Methods 0.000 description 1
- 210000002288 golgi apparatus Anatomy 0.000 description 1
- 101150028578 grp78 gene Proteins 0.000 description 1
- 229940029575 guanosine Drugs 0.000 description 1
- 239000000118 hair dye Substances 0.000 description 1
- 125000001475 halogen functional group Chemical group 0.000 description 1
- 108010002430 hemicellulase Proteins 0.000 description 1
- 108010018734 hexose oxidase Proteins 0.000 description 1
- 238000004128 high performance liquid chromatography Methods 0.000 description 1
- 229940088597 hormone Drugs 0.000 description 1
- 239000005556 hormone Substances 0.000 description 1
- 238000009396 hybridization Methods 0.000 description 1
- 108010002685 hygromycin-B kinase Proteins 0.000 description 1
- 210000001822 immobilized cell Anatomy 0.000 description 1
- 230000003100 immobilizing effect Effects 0.000 description 1
- 210000000987 immune system Anatomy 0.000 description 1
- 230000003053 immunization Effects 0.000 description 1
- 238000002649 immunization Methods 0.000 description 1
- 238000003018 immunoassay Methods 0.000 description 1
- 230000002163 immunogen Effects 0.000 description 1
- 229940027941 immunoglobulin g Drugs 0.000 description 1
- 238000000099 in vitro assay Methods 0.000 description 1
- 238000010874 in vitro model Methods 0.000 description 1
- 238000011065 in-situ storage Methods 0.000 description 1
- 239000004615 ingredient Substances 0.000 description 1
- 239000003112 inhibitor Substances 0.000 description 1
- 230000000977 initiatory effect Effects 0.000 description 1
- CDAISMWEOUEBRE-GPIVLXJGSA-N inositol Chemical compound O[C@H]1[C@H](O)[C@@H](O)[C@H](O)[C@H](O)[C@@H]1O CDAISMWEOUEBRE-GPIVLXJGSA-N 0.000 description 1
- 229960000367 inositol Drugs 0.000 description 1
- 229940125396 insulin Drugs 0.000 description 1
- 229940047124 interferons Drugs 0.000 description 1
- 239000001573 invertase Substances 0.000 description 1
- 235000011073 invertase Nutrition 0.000 description 1
- 238000005342 ion exchange Methods 0.000 description 1
- 238000004255 ion exchange chromatography Methods 0.000 description 1
- 239000012948 isocyanate Substances 0.000 description 1
- 150000002513 isocyanates Chemical class 0.000 description 1
- 150000002540 isothiocyanates Chemical class 0.000 description 1
- 229960000318 kanamycin Drugs 0.000 description 1
- SBUJHOSQTJFQJX-NOAMYHISSA-N kanamycin Chemical compound O[C@@H]1[C@@H](O)[C@H](O)[C@@H](CN)O[C@@H]1O[C@H]1[C@H](O)[C@@H](O[C@@H]2[C@@H]([C@@H](N)[C@H](O)[C@@H](CO)O2)O)[C@H](N)C[C@@H]1N SBUJHOSQTJFQJX-NOAMYHISSA-N 0.000 description 1
- 229930027917 kanamycin Natural products 0.000 description 1
- 229930182823 kanamycin A Natural products 0.000 description 1
- 210000003734 kidney Anatomy 0.000 description 1
- 229940116108 lactase Drugs 0.000 description 1
- FCCDDURTIIUXBY-UHFFFAOYSA-N lipoamide Chemical compound NC(=O)CCCCC1CCSS1 FCCDDURTIIUXBY-UHFFFAOYSA-N 0.000 description 1
- 239000006210 lotion Substances 0.000 description 1
- 238000004020 luminiscence type Methods 0.000 description 1
- 229940040129 luteinizing hormone Drugs 0.000 description 1
- 101150039489 lysZ gene Proteins 0.000 description 1
- 239000006166 lysate Substances 0.000 description 1
- 229960000274 lysozyme Drugs 0.000 description 1
- 239000004325 lysozyme Substances 0.000 description 1
- 235000010335 lysozyme Nutrition 0.000 description 1
- 230000014759 maintenance of location Effects 0.000 description 1
- 238000013507 mapping Methods 0.000 description 1
- 230000007246 mechanism Effects 0.000 description 1
- 230000001404 mediated effect Effects 0.000 description 1
- 239000002207 metabolite Substances 0.000 description 1
- 229910052751 metal Inorganic materials 0.000 description 1
- 239000002184 metal Substances 0.000 description 1
- 229910021645 metal ion Inorganic materials 0.000 description 1
- 125000002496 methyl group Chemical group [H]C([H])([H])* 0.000 description 1
- 108010009355 microbial metalloproteinases Proteins 0.000 description 1
- 244000000010 microbial pathogen Species 0.000 description 1
- 239000003226 mitogen Substances 0.000 description 1
- 230000011278 mitosis Effects 0.000 description 1
- 102000035118 modified proteins Human genes 0.000 description 1
- 108091005573 modified proteins Proteins 0.000 description 1
- 239000003068 molecular probe Substances 0.000 description 1
- 229910000402 monopotassium phosphate Inorganic materials 0.000 description 1
- 235000019796 monopotassium phosphate Nutrition 0.000 description 1
- 229960004927 neomycin Drugs 0.000 description 1
- 101150095344 niaD gene Proteins 0.000 description 1
- 229910052759 nickel Inorganic materials 0.000 description 1
- MGFYIUFZLHCRTH-UHFFFAOYSA-N nitrilotriacetic acid Chemical compound OC(=O)CN(CC(O)=O)CC(O)=O MGFYIUFZLHCRTH-UHFFFAOYSA-N 0.000 description 1
- LQNUZADURLCDLV-UHFFFAOYSA-N nitrobenzene Substances [O-][N+](=O)C1=CC=CC=C1 LQNUZADURLCDLV-UHFFFAOYSA-N 0.000 description 1
- 101150105920 npr gene Proteins 0.000 description 1
- 101150017837 nprM gene Proteins 0.000 description 1
- 238000007899 nucleic acid hybridization Methods 0.000 description 1
- 239000012038 nucleophile Substances 0.000 description 1
- 229960002446 octanoic acid Drugs 0.000 description 1
- 239000002751 oligonucleotide probe Substances 0.000 description 1
- 238000002515 oligonucleotide synthesis Methods 0.000 description 1
- 108090000021 oryzin Proteins 0.000 description 1
- 210000001672 ovary Anatomy 0.000 description 1
- 150000002924 oxiranes Chemical class 0.000 description 1
- 229940055729 papain Drugs 0.000 description 1
- 235000019834 papain Nutrition 0.000 description 1
- 239000000199 parathyroid hormone Substances 0.000 description 1
- 229960001319 parathyroid hormone Drugs 0.000 description 1
- 230000036961 partial effect Effects 0.000 description 1
- 239000013618 particulate matter Substances 0.000 description 1
- 108010087558 pectate lyase Proteins 0.000 description 1
- 239000001814 pectin Substances 0.000 description 1
- 229920001277 pectin Polymers 0.000 description 1
- 235000010987 pectin Nutrition 0.000 description 1
- 101150019841 penP gene Proteins 0.000 description 1
- 239000000813 peptide hormone Substances 0.000 description 1
- 239000000137 peptide hydrolase inhibitor Substances 0.000 description 1
- 210000001322 periplasm Anatomy 0.000 description 1
- JTJMJGYZQZDUJJ-UHFFFAOYSA-N phencyclidine Chemical compound C1CCCCN1C1(C=2C=CC=CC=2)CCCCC1 JTJMJGYZQZDUJJ-UHFFFAOYSA-N 0.000 description 1
- 239000010452 phosphate Substances 0.000 description 1
- 108010082527 phosphinothricin N-acetyltransferase Proteins 0.000 description 1
- 239000011574 phosphorus Substances 0.000 description 1
- 229910052698 phosphorus Inorganic materials 0.000 description 1
- 229940085127 phytase Drugs 0.000 description 1
- 235000002949 phytic acid Nutrition 0.000 description 1
- 229940127126 plasminogen activator Drugs 0.000 description 1
- 229920003023 plastic Polymers 0.000 description 1
- 239000004033 plastic Substances 0.000 description 1
- 229920002401 polyacrylamide Polymers 0.000 description 1
- 229920000573 polyethylene Polymers 0.000 description 1
- 210000004896 polypeptide structure Anatomy 0.000 description 1
- 230000001323 posttranslational effect Effects 0.000 description 1
- GNSKLFRGEWLPPA-UHFFFAOYSA-M potassium dihydrogen phosphate Chemical compound [K+].OP(O)([O-])=O GNSKLFRGEWLPPA-UHFFFAOYSA-M 0.000 description 1
- OXCMYAYHXIHQOA-UHFFFAOYSA-N potassium;[2-butyl-5-chloro-3-[[4-[2-(1,2,4-triaza-3-azanidacyclopenta-1,4-dien-5-yl)phenyl]phenyl]methyl]imidazol-4-yl]methanol Chemical compound [K+].CCCCC1=NC(Cl)=C(CO)N1CC1=CC=C(C=2C(=CC=CC=2)C2=N[N-]N=N2)C=C1 OXCMYAYHXIHQOA-UHFFFAOYSA-N 0.000 description 1
- 230000001376 precipitating effect Effects 0.000 description 1
- 238000001556 precipitation Methods 0.000 description 1
- 239000002243 precursor Substances 0.000 description 1
- 230000002265 prevention Effects 0.000 description 1
- 150000003141 primary amines Chemical class 0.000 description 1
- 229940097325 prolactin Drugs 0.000 description 1
- 230000035755 proliferation Effects 0.000 description 1
- 235000013930 proline Nutrition 0.000 description 1
- 229960002429 proline Drugs 0.000 description 1
- 108010017378 prolyl aminopeptidase Proteins 0.000 description 1
- 235000019260 propionic acid Nutrition 0.000 description 1
- OANVFVBYPNXRLD-UHFFFAOYSA-M propyromazine bromide Chemical compound [Br-].C12=CC=CC=C2SC2=CC=CC=C2N1C(=O)C(C)[N+]1(C)CCCC1 OANVFVBYPNXRLD-UHFFFAOYSA-M 0.000 description 1
- 238000001742 protein purification Methods 0.000 description 1
- 238000005086 pumping Methods 0.000 description 1
- 238000003908 quality control method Methods 0.000 description 1
- IUVKMZGDUIUOCP-BTNSXGMBSA-N quinbolone Chemical compound O([C@H]1CC[C@H]2[C@H]3[C@@H]([C@]4(C=CC(=O)C=C4CC3)C)CC[C@@]21C)C1=CCCC1 IUVKMZGDUIUOCP-BTNSXGMBSA-N 0.000 description 1
- 239000011535 reaction buffer Substances 0.000 description 1
- 239000012429 reaction media Substances 0.000 description 1
- 239000011541 reaction mixture Substances 0.000 description 1
- 230000009257 reactivity Effects 0.000 description 1
- 230000008929 regeneration Effects 0.000 description 1
- 238000011069 regeneration method Methods 0.000 description 1
- 239000003488 releasing hormone Substances 0.000 description 1
- 230000003362 replicative effect Effects 0.000 description 1
- 230000000717 retained effect Effects 0.000 description 1
- 230000002441 reversible effect Effects 0.000 description 1
- 238000012552 review Methods 0.000 description 1
- 239000002342 ribonucleoside Substances 0.000 description 1
- 239000002336 ribonucleotide Substances 0.000 description 1
- 125000002652 ribonucleotide group Chemical group 0.000 description 1
- 101150025220 sacB gene Proteins 0.000 description 1
- 230000007017 scission Effects 0.000 description 1
- CDAISMWEOUEBRE-UHFFFAOYSA-N scyllo-inosotol Natural products OC1C(O)C(O)C(O)C(O)C1O CDAISMWEOUEBRE-UHFFFAOYSA-N 0.000 description 1
- 238000007789 sealing Methods 0.000 description 1
- 210000004739 secretory vesicle Anatomy 0.000 description 1
- 238000012163 sequencing technique Methods 0.000 description 1
- 239000002453 shampoo Substances 0.000 description 1
- 238000003530 single readout Methods 0.000 description 1
- 238000002741 site-directed mutagenesis Methods 0.000 description 1
- 239000002884 skin cream Substances 0.000 description 1
- 150000003384 small molecules Chemical class 0.000 description 1
- AWUCVROLDVIAJX-GSVOUGTGSA-N sn-glycerol 3-phosphate Chemical compound OC[C@@H](O)COP(O)(O)=O AWUCVROLDVIAJX-GSVOUGTGSA-N 0.000 description 1
- 239000000344 soap Substances 0.000 description 1
- 235000017550 sodium carbonate Nutrition 0.000 description 1
- 229910000029 sodium carbonate Inorganic materials 0.000 description 1
- 239000011780 sodium chloride Substances 0.000 description 1
- 238000002415 sodium dodecyl sulfate polyacrylamide gel electrophoresis Methods 0.000 description 1
- 239000002904 solvent Substances 0.000 description 1
- 210000001082 somatic cell Anatomy 0.000 description 1
- NHXLMOGPVYXJNR-ATOGVRKGSA-N somatostatin Chemical compound C([C@H]1C(=O)N[C@H](C(N[C@@H](CO)C(=O)N[C@@H](CSSC[C@@H](C(=O)N[C@@H](CCCCN)C(=O)N[C@@H](CC(N)=O)C(=O)N[C@@H](CC=2C=CC=CC=2)C(=O)N[C@@H](CC=2C=CC=CC=2)C(=O)N[C@@H](CC=2C3=CC=CC=C3NC=2)C(=O)N[C@@H](CCCCN)C(=O)N[C@H](C(=O)N1)[C@@H](C)O)NC(=O)CNC(=O)[C@H](C)N)C(O)=O)=O)[C@H](O)C)C1=CC=CC=C1 NHXLMOGPVYXJNR-ATOGVRKGSA-N 0.000 description 1
- 229960000553 somatostatin Drugs 0.000 description 1
- 230000009870 specific binding Effects 0.000 description 1
- 235000019698 starch Nutrition 0.000 description 1
- 230000000638 stimulation Effects 0.000 description 1
- 238000003756 stirring Methods 0.000 description 1
- 239000011550 stock solution Substances 0.000 description 1
- 229960005202 streptokinase Drugs 0.000 description 1
- 238000007920 subcutaneous administration Methods 0.000 description 1
- KDYFGRWQOYBRFD-UHFFFAOYSA-L succinate(2-) Chemical compound [O-]C(=O)CCC([O-])=O KDYFGRWQOYBRFD-UHFFFAOYSA-L 0.000 description 1
- 229940014800 succinic anhydride Drugs 0.000 description 1
- 239000005720 sucrose Substances 0.000 description 1
- 150000003871 sulfonates Chemical class 0.000 description 1
- YBBRCQOCSYXUOC-UHFFFAOYSA-N sulfuryl dichloride Chemical class ClS(Cl)(=O)=O YBBRCQOCSYXUOC-UHFFFAOYSA-N 0.000 description 1
- 235000011149 sulphuric acid Nutrition 0.000 description 1
- 239000013589 supplement Substances 0.000 description 1
- 239000004094 surface-active agent Substances 0.000 description 1
- 108010075550 termamyl Proteins 0.000 description 1
- 229960002180 tetracycline Drugs 0.000 description 1
- 229930101283 tetracycline Natural products 0.000 description 1
- 235000019364 tetracycline Nutrition 0.000 description 1
- 150000003522 tetracyclines Chemical class 0.000 description 1
- 150000003573 thiols Chemical class 0.000 description 1
- 150000003585 thioureas Chemical class 0.000 description 1
- LCJVIYPJPCBWKS-NXPQJCNCSA-N thymosin Chemical compound SC[C@@H](N)C(=O)N[C@H](CO)C(=O)N[C@H](CC(O)=O)C(=O)N[C@@H](C)C(=O)N[C@@H](C)C(=O)N[C@H](C(C)C)C(=O)N[C@H](CC(O)=O)C(=O)N[C@H](C(C)C)C(=O)N[C@H](CO)C(=O)N[C@H](CO)C(=O)N[C@H](CCC(O)=O)C(=O)N[C@H]([C@@H](C)CC)C(=O)N[C@H]([C@H](C)O)C(=O)N[C@H](C(C)C)C(=O)N[C@H](CCCCN)C(=O)N[C@H](CC(O)=O)C(=O)N[C@H](CC(C)C)C(=O)N[C@H](CCCCN)C(=O)N[C@H](CCC(O)=O)C(=O)N[C@H](CCCCN)C(=O)N[C@H](CCCCN)C(=O)N[C@H](CCC(O)=O)C(=O)N[C@H](C(C)C)C(=O)N[C@H](C(C)C)C(=O)N[C@H](CCC(O)=O)C(=O)N[C@H](CCC(O)=O)C(=O)N[C@@H](C)C(=O)N[C@H](CCC(O)=O)C(O)=O LCJVIYPJPCBWKS-NXPQJCNCSA-N 0.000 description 1
- 229960000187 tissue plasminogen activator Drugs 0.000 description 1
- JOXIMZWYDAKGHI-UHFFFAOYSA-M toluene-4-sulfonate Chemical compound CC1=CC=C(S([O-])(=O)=O)C=C1 JOXIMZWYDAKGHI-UHFFFAOYSA-M 0.000 description 1
- 229940034610 toothpaste Drugs 0.000 description 1
- 239000000606 toothpaste Substances 0.000 description 1
- 125000005490 tosylate group Chemical group 0.000 description 1
- 101150080369 tpiA gene Proteins 0.000 description 1
- 230000005030 transcription termination Effects 0.000 description 1
- 238000012546 transfer Methods 0.000 description 1
- 238000006276 transfer reaction Methods 0.000 description 1
- QORWJWZARLRLPR-UHFFFAOYSA-H tricalcium bis(phosphate) Chemical compound [Ca+2].[Ca+2].[Ca+2].[O-]P([O-])([O-])=O.[O-]P([O-])([O-])=O QORWJWZARLRLPR-UHFFFAOYSA-H 0.000 description 1
- LENZDBCJOHFCAS-UHFFFAOYSA-N tris Chemical compound OCC(N)(CO)CO LENZDBCJOHFCAS-UHFFFAOYSA-N 0.000 description 1
- 101150016309 trpC gene Proteins 0.000 description 1
- 241000701161 unidentified adenovirus Species 0.000 description 1
- DRTQHJPVMGBUCF-UHFFFAOYSA-N uracil arabinoside Natural products OC1C(O)C(CO)OC1N1C(=O)NC(=O)C=C1 DRTQHJPVMGBUCF-UHFFFAOYSA-N 0.000 description 1
- 150000003672 ureas Chemical class 0.000 description 1
- 229940045145 uridine Drugs 0.000 description 1
- 238000003828 vacuum filtration Methods 0.000 description 1
- 230000003612 virological effect Effects 0.000 description 1
- 239000011534 wash buffer Substances 0.000 description 1
- XLYOFNOQVPJJNP-UHFFFAOYSA-N water Substances O XLYOFNOQVPJJNP-UHFFFAOYSA-N 0.000 description 1
- 210000002268 wool Anatomy 0.000 description 1
- 101150110790 xylB gene Proteins 0.000 description 1
- 229920001221 xylan Polymers 0.000 description 1
- 150000004823 xylans Chemical class 0.000 description 1
- 108010078692 yeast proteinase B Proteins 0.000 description 1
Classifications
-
- C—CHEMISTRY; METALLURGY
- C40—COMBINATORIAL TECHNOLOGY
- C40B—COMBINATORIAL CHEMISTRY; LIBRARIES, e.g. CHEMICAL LIBRARIES
- C40B30/00—Methods of screening libraries
- C40B30/04—Methods of screening libraries by measuring the ability to specifically bind a target molecule, e.g. antibody-antigen binding, receptor-ligand binding
-
- G—PHYSICS
- G01—MEASURING; TESTING
- G01N—INVESTIGATING OR ANALYSING MATERIALS BY DETERMINING THEIR CHEMICAL OR PHYSICAL PROPERTIES
- G01N33/00—Investigating or analysing materials by specific methods not covered by groups G01N1/00 - G01N31/00
- G01N33/48—Biological material, e.g. blood, urine; Haemocytometers
- G01N33/50—Chemical analysis of biological material, e.g. blood, urine; Testing involving biospecific ligand binding methods; Immunological testing
- G01N33/68—Chemical analysis of biological material, e.g. blood, urine; Testing involving biospecific ligand binding methods; Immunological testing involving proteins, peptides or amino acids
- G01N33/6803—General methods of protein analysis not limited to specific proteins or families of proteins
-
- G—PHYSICS
- G01—MEASURING; TESTING
- G01N—INVESTIGATING OR ANALYSING MATERIALS BY DETERMINING THEIR CHEMICAL OR PHYSICAL PROPERTIES
- G01N33/00—Investigating or analysing materials by specific methods not covered by groups G01N1/00 - G01N31/00
- G01N33/48—Biological material, e.g. blood, urine; Haemocytometers
- G01N33/50—Chemical analysis of biological material, e.g. blood, urine; Testing involving biospecific ligand binding methods; Immunological testing
- G01N33/68—Chemical analysis of biological material, e.g. blood, urine; Testing involving biospecific ligand binding methods; Immunological testing involving proteins, peptides or amino acids
- G01N33/6803—General methods of protein analysis not limited to specific proteins or families of proteins
- G01N33/6842—Proteomic analysis of subsets of protein mixtures with reduced complexity, e.g. membrane proteins, phosphoproteins, organelle proteins
-
- G—PHYSICS
- G01—MEASURING; TESTING
- G01N—INVESTIGATING OR ANALYSING MATERIALS BY DETERMINING THEIR CHEMICAL OR PHYSICAL PROPERTIES
- G01N33/00—Investigating or analysing materials by specific methods not covered by groups G01N1/00 - G01N31/00
- G01N33/48—Biological material, e.g. blood, urine; Haemocytometers
- G01N33/50—Chemical analysis of biological material, e.g. blood, urine; Testing involving biospecific ligand binding methods; Immunological testing
- G01N33/68—Chemical analysis of biological material, e.g. blood, urine; Testing involving biospecific ligand binding methods; Immunological testing involving proteins, peptides or amino acids
- G01N33/6803—General methods of protein analysis not limited to specific proteins or families of proteins
- G01N33/6845—Methods of identifying protein-protein interactions in protein mixtures
Definitions
- This invention relates to methods for screening libraries of protein variants for those having reduced immunogenicity as compared to a protein backbone. More specifically, the present invention provides a method for identifying protein variants with reduced immunogenicity in an efficient manner.
- proteins including enzymes
- allergic response allergy, allergenic and allergenicity are used according to their usual definitions, i.e. to describe the reaction due to immune responses wherein the antibody most often is IgE, less often IgG4 and diseases due to this immune response.
- Allergic diseases include urticaria, hay-fever, asthma, and atopic dermatitis. They may even evolve into an anaphylactic shock.
- the sensitization phase involves a first exposure of an individual to an allergen. This event activates specific T- and B-lymphocytes, and leads to the production of allergen-specific IgE antibodies (in the present context the antibodies are denoted as usual, i.e. immunoglobulin E is IgE etc.). These IgE antibodies eventually facilitate allergen capturing and presentation to T-lymphocytes at the onset of the symptomatic phase.
- This phase is initiated by a second exposure to the same or a resembling antigen.
- the specific IgE antibodies bind to the specific IgE receptors on mast cells and basophils, among others, in a mode that allows antigen binding to the cell-bound IgE antibodies.
- the polyclonal nature of this process results in bridging and clustering of the IgE receptors, and subsequently in the activation of mast cells and basophiles. This activation triggers the release of various chemical mediators involved in the early as well as late phase reactions of the symptomatic phase of allergy.
- HTS methods are designed to detect an improved functionality of the expressed protein variants.
- variants with debilitated functionality which are also likely to bind antibodies poorly
- To overcome this limitation it is necessary to deviate from the design of most present HTS methods and introduce a dual assay system in which both antibody binding capacity and functionality are determined without compromising the high throughput capability necessary to benefit the diversity of diversified libraries.
- the prior art, different publications have disclosed aspects that are useful for creating protein variants with altered antibody binding capacity, but none have disclosed a method for screening a diversified library expressed in host cells to search for functional protein variants
- WO 99/447680 discloses the modification of B-cell epitopes by protein engineering.
- the method is based on crystal structures of Fab-antigen complexes, and B-cell epitopes are defined as "a section of the surface of the antigen comprising 15-25 amino acid residues, which are within a distance from the atoms of the antibody enabling direct interaction" (p.3).
- This publication does not show how one selects which Fab fragment to use (e.g. to target the most dominant allergy epitopes) or how one selects the substitutions to be made.
- their method cannot be used in the absence of such crystallographic data, which is very cumbersome, sometimes impossible, to obtain - especially since one would need a separate crystal structure for each epitope to be changed.
- Slootstra et al; Molecular Diversity, 2, pp. 156-164, 1996 disclose the screening of a semirandom library of peptides for their binding properties to three monoclonal antibodies by immobilizing the peptides on polyethylene pins and binding a dilution series of each antibody to the pins.
- all peptides are prepared by chemical synthesis; hence, it does not disclose any method to overcome background problems from using gene-encoded polypeptides expressed in a microbial host.
- the antibody binding assays are based on 10-step dilution series for each antibody, meaning that 30 separate assays are necessary to evaluate each test compound. This makes the disclosed methods insufficient for high- throughput screening.
- WO 97/30150 discloses the construction and expression of diversified libraries of a myelin basic protein (MBP) and the analysis of these variants by testing their T-cell antagonizing activity.
- MBP myelin basic protein
- This reference only tests for 'trans-dominant effects' (p. 17) in which a single peptide harboring a productive mutation will show up even in the presence of 10 unchanged peptide fragments.
- This means that dysfunctional or poorly expressed protein variants will show no response and hence, that the teachings of this reference cannot be used for devising an assay to identify protein variants with reduced antibody binding capacity and retained functionality.
- the problem to be solved by the present invention is to provide a method to perform assays that efficiently and accurately can screen large numbers of cell populations producing variants of a molecule of interest.
- the invention relates to a method for high throughput screening (HTS) of a large population of host cells for production of a molecule of interest.
- HTS high throughput screening
- the invention relates to a method for screening a library of protein variants for functional variants with reduced antibody binding capacity, comprising the steps of:
- the term “isolated” indicates that the protein is found in a condition other than its native environment, such as apart from blood and animal tissue. In a prefened form, the isolated protein is substantially free of other proteins, particularly other proteins of 5 animal origin. It is prefened to provide the proteins in a highly purified form, i.e., greater than 95% pure, more preferably greater than 99% pure.
- the term “isolated” indicates that the molecule is removed from its natural genetic milieu, and is thus free of other extraneous or unwanted coding sequences, and is in a form suitable for use within genetically engineered protein production systems.
- isolated o molecules are those that are separated from their natural environment and include cDNA and genomic clones.
- isolated DNA molecules of the present invention are free of other genes with which they are ordinarily associated, and may include naturally occurring 5' and 3' untranslated regions such as promoters and terminators. The identification of associated regions will be evident to one of ordinary skill in the art (see for example, Dynan and Tijan, Nature 316: 774-78, 1985).
- polynucleotide is a single- or double-stranded polymer of deoxyribonucleotide or ribonucleotide bases read from the 5' to the 3' end.
- Polynucleotides include RNA and DNA, and may be isolated from natural sources, synthesized in vitro, or prepared from a combination of natural and synthetic molecules.
- nucleic acid molecule refers to the phosphate ester polymeric form of ribonucleosides (adenosine, guanosine, uridine or cytidine; "RNA molecules”) or deoxyribonucleosides (deoxyadenosine, deoxyguanosine, deoxythymidine, or deoxycytidine; "DNA molecules”) in either single stranded form, or a double-stranded helix. Double stranded DNA-DNA, DNA- RNA and RNA-RNA helices are possible.
- nucleic acid molecule refers only to the primary and secondary structure of the molecule, and does not limit it to any particular tertiary or quaternary forms.
- this term includes double-stranded DNA found, inter alia, in linear or circular DNA molecules (e.g., restriction fragments), plasmids, and chromosomes.
- sequences may be described herein according to the normal convention of giving only the sequence in the 5' to 3' direction along the nontranscribed strand of DNA (i.e., the strand having a sequence homologous to the mRNA).
- a "recombinant DNA molecule” is a DNA molecule that has undergone a molecular biological manipulation.
- a DNA "coding sequence” is a double-stranded DNA sequence, which is transcribed and translated into a polypeptide in a cell in vitro or in vivo when placed under the control of appropriate regulatory sequences. The boundaries of the coding sequence are determined by a start codon at the 5' (amino) terminus and a translation stop codon at the 3' (carboxyl) terminus.
- a coding sequence can include, but is not limited to, prokaryotic sequences, cDNA from eukaryotic mRNA, genomic DNA sequences from eukaryotic (e.g., mammalian) DNA, and even synthetic DNA sequences. If the coding sequence is intended for expression in a eukaryotic cell, a polyadenylation signal and transcription termination sequence will usually be located 3' to the coding sequence.
- An “Expression vector” is a DNA molecule, linear or circular, that comprises a segment encoding a polypeptide of interest operably linked to additional segments that provide for its transcription. Such additional segments may include promoter and terminator sequences, and optionally one or more origins of replication, one or more selectable markers, an enhancer, a polyadenylation signal, and the like. Expression vectors are generally derived from plasmid or viral DNA, or may contain elements of both.
- Transcriptional and translational control sequences are DNA regulatory sequences, such as promoters, enhancers, terminators, and the like, that provide for the expression of a coding sequence in a host cell.
- polyadenylation signals are control sequences.
- a “secretory signal sequence” is a DNA sequence that encodes a polypeptide (a "secretory peptide” that, as a component of a larger polypeptide, directs the larger polypeptide through a secretory pathway of a cell in which it is synthesized.
- the larger polypeptide is commonly cleaved to remove the secretory peptide during transit through the secretory pathway.
- promoter is used herein for its art-recognized meaning to denote a portion of a gene containing DNA sequences that provide for the binding of RNA polymerase and initiation of transcription. Promoter sequences are commonly, but not always, found in the 5' non-coding regions of genes.
- “Operably linked”, when referring to DNA segments, indicates that the segments are ananged so that they function in concert for their intended purposes, e.g. transcription initiates in the promoter and proceeds through the coding segment to the terminator.
- a coding sequence is "under the control" of transcriptional and translational control sequences in a cell when RNA polymerase transcribes the coding sequence into mRNA, which is then trans-RNA spliced and translated into the protein encoded by the coding sequence.
- isolated polypeptide is a polypeptide which is essentially free of other non-[enzyme] polypeptides, e.g., at least about 20% pure, preferably at least about 40% pure, more preferably about 60% pure, even more preferably about 80% pure, most preferably about 90% pure, and even most preferably about 95% pure, as determined by SDS-PAGE.
- Heterologous DNA refers to DNA not naturally located in the cell, or in a chromosomal site of the cell.
- the heterologous DNA includes a gene foreign to the cell.
- a cell has been "transfected” by exogenous or heterologous DNA when such DNA has been introduced inside the cell.
- a cell has been "transformed” by exogenous or heterologous DNA when the transfected DNA effects a phenotypic change.
- the transforming DNA should be integrated (covalently linked) into chromosomal DNA making up the genome of the cell.
- a "clone” is a population of cells derived from a single cell or common ancestor by mitosis.
- Homologous recombination refers to the insertion of a foneign DNA sequence of a vector in a chromosome.
- the vector targets a specific chromosomal site for homologous recombination.
- the vector will contain sufficiently long regions of homology to sequences of the chromosome to allow complementary binding and incorporation of the vector into the chromosome. Longer regions of homology, and greater degrees of sequence similarity, may increase the efficiency of homologous recombination.
- the techniques used to isolate or clone a nucleic acid sequence encoding a polypeptide include isolation from genomic DNA, preparation from cDNA, or a combination thereof.
- the cloning of the nucleic acid sequences of the present invention from such genomic DNA can be effected, e.g., by using the well known polymerase chain reaction (PCR) or antibody screening of expression libraries to detect cloned DNA fragments with shared structural features. See, e.g., Innis et al., 1990, A Guide to Methods and Application, Academic Press, New York.
- nucleic acid amplification procedures such as ligase chain reaction (LCR), ligated activated transcription (LAT) and nuceic acid sequence-based amplification (NASBA) may be used.
- LCR ligase chain reaction
- LAT ligated activated transcription
- NASBA nuceic acid sequence-based amplification
- the nucleic acid sequence may be cloned from a strain producing the polypeptide, or from another related organism and thus, for example, may be an allelic or species variant of the polypeptide encoding region of the nucleic acid sequence.
- isolated nucleic acid sequence refers to a nucleic acid sequence which is essentially free of other nucleic acid sequences, e.g., at least about 20% pure, preferably at least about 40% pure, more preferably about 60% pure, even more preferably about 80% pure, most preferably about 90% pure, and even most preferably about 95% pure, as determined by agarose gel electorphoresis.
- an isolated nucleic acid sequence can be obtained by standard cloning procedures used in genetic engineering to relocate the nucleic acid sequence from its natural location to a different site where it will be reproduced.
- the cloning procedures may involve excision and isolation of a desired nucleic acid fragment comprising the nucleic acid sequence encoding the polypeptide, insertion of the fragment into a vector molecule, and incorporation of the recombinant vector into a host cell where multiple copies or clones of the nucleic acid sequence will be replicated.
- the nucleic acid sequence may be of genomic, cDNA, RNA, semisynthetic, synthetic origin, or any combinations thereof.
- nucleic acid construct is intended to indicate any nucleic acid molecule of cDNA, genomic DNA, synthetic DNA or RNA origin.
- construct is intended to indicate a nucleic acid segment which may be single- or double-stranded, and which may be based on a complete or partial naturally occurring nucleotide sequence encoding a polypeptide of interest.
- the construct may optionally contain other nucleic acid segments.
- the DNA of interest may suitably be of genomic or cDNA origin, for instance obtained by preparing a genomic or cDNA library and screening for DNA sequences coding for all or part of the polypeptide by hybridization using synthetic oligonucleotide probes in accordance with standard techniques (cf. Sambrook et al., supra).
- the nucleic acid construct may also be prepared synthetically by established standard methods, e.g. the phosphoamidite method described by Beaucage and Caruthers, Tetrahedron Letters 22 (1981), 1859 - 1869, or the method described by Matthes et al, EMBO Journal 3 (1984), 801 - 805.
- oligonucleotides are synthesized, 5 e.g. in an automatic DNA synthesizer, purified, annealed, ligated and cloned in suitable vectors.
- the nucleic acid construct may be of mixed synthetic and genomic, mixed synthetic and cDNA or mixed genomic and cDNA origin prepared by ligating fragments of o synthetic, genomic or cDNA origin (as appropriate), the fragments conesponding to various parts of the entire nucleic acid construct, in accordance with standard techniques.
- the nucleic acid construct may also be prepared by polymerase chain reaction using specific primers, for instance as described in US 4,683,202 or Saiki et al., Science 239 (1988), 487 - s 491.
- nucleic acid construct may be synonymous with the term expression cassette when the nucleic acid construct contains all the control sequences required for expression of a coding sequence of the present invention.
- coding sequence as defined herein is a o sequence which is transcribed into mRNA and translated into a polypeptide of the present invention when placed under the control of the above mentioned control sequences. The boundaries of the coding sequence are generally determined by a translation start codon ATG at the 5 '-terminus and a translation stop codon at the 3 '-terminus.
- a coding sequence can include, but is not limited to, DNA, cDNA, and recombinant nucleic acid sequences. 5
- control sequences is defined herein to include all components which are necessary or advantageous for expression of the coding sequence of the nucleic acid sequence.
- Each control sequence may be native or foreign to the nucleic acid sequence encoding the polypeptide.
- control sequences include, but are not limited to, a leader, a 0 polyadenylation sequence, a propeptide sequence, a promoter, a signal sequence, and a transcription terminator.
- the control sequences include a promoter, and transcriptional and translational stop signals.
- the control sequences may be provided with linkers for the purpose of introducing specific restriction sites facilitating ligation of the control sequences with the coding region of the nucleic acid sequence encoding a polypeptide.
- the control sequence may be an appropriate promoter sequence, a nucleic acid sequence which is recognized by a host cell for expression of the nucleic acid sequence.
- the promoter sequence contains transcription and translation control sequences which mediate the expression of the polypeptide.
- the promoter may be any nucleic acid sequence which shows transcriptional activity in the host cell of choice and may be obtained from genes encoding extracellular or intracellular polypeptides either homologous or heterologous to the host cell.
- the control sequence may also be a suitable transcription terminator sequence, a sequence recognized by a host cell to terminate transcription. The terminator sequence is operably linked to the 3' terminus of the nucleic acid sequence encoding the polypeptide. Any terminator which is functional in the host cell of choice may be used in the present invention.
- the control sequence may also be a polyadenylation sequence, a sequence which is operably linked to the 3' terminus of the nucleic acid sequence and which, when transcribed, is recognized by the host cell as a signal to add polyadenosine residues to transcribed mRNA. Any polyadenylation sequence which is functional in the host cell of choice may be used in the present invention.
- the control sequence may also be a signal peptide coding region, which codes for an amino acid sequence linked to the amino terminus of the polypeptide which can direct the expressed polypeptide into the cell's secretory pathway of the host cell.
- the 5' end of the coding sequence of the nucleic acid sequence may inherently contain a signal peptide coding region naturally linked in translation reading frame with the segment of the coding region which encodes the secreted polypeptide.
- the 5' end of the coding sequence may contain a signal peptide coding region which is foreign to that portion of the coding sequence which encodes the secreted polypeptide.
- a foreign signal peptide coding region may be required where the coding sequence does not normally contain a signal peptide coding region.
- the foreign signal peptide coding region may simply replace the natural signal peptide coding region in order to obtain enhanced secretion relative to the natural signal peptide coding region normally associated with the coding sequence.
- the signal peptide coding region may be obtained from a glucoamylase or an amylase gene from an Aspergillus species, a lipase or proteinase gene from a Rhizomucor species, the gene for the alpha-factor from Saccharomyces cerevisiae, an amylase or a protease gene from a Bacillus species, or the calf preprochymosin gene.
- any signal peptide coding region capable of directing the expressed polypeptide into the secretory pathway of a host cell of choice may be used in the present invention.
- the control sequence may also be a propeptide coding region, which codes for an amino acid sequence positioned at the amino terminus of a polypeptide.
- the resultant polypeptide is known as a proenzyme or propolypeptide (or a zymogen in some cases).
- a propolypeptide is generally inactive and can be converted to mature active polypeptide by catalytic or autocatalytic cleavage of the propeptide from the propolypeptide.
- the propeptide coding region may be obtained from the Bacillus subtilis alkaline protease gene (aprE), the Bacillus subtilis neutral protease gene (nprT), the Saccharomyces cerevisiae alpha-factor gene, or the Myceliophthora thermophilum laccase gene (WO 95/33836).
- the nucleic acid constructs of the present invention may also comprise one or more nucleic acid sequences which encode one or more factors that are advantageous in the expression of the polypeptide, e.g., an activator (e.g., a trans-acting factor), a chaperone, and a processing protease. Any factor that is functional in the host cell of choice may be used in the present invention.
- an activator e.g., a trans-acting factor
- a chaperone e.g., a chaperone
- processing protease e.g., a factor that is functional in the host cell of choice.
- the nucleic acids encoding one or more of these factors are not necessarily in tandem with the nucleic acid sequence encoding the polypeptide.
- An activator is a protein which activates transcription of a nucleic acid sequence encoding a polypeptide (Kudla et al., 1990, EMBO Journal 9:1355-1364; Jarai and Buxton, 1994, Cunent Genetics 26:2238-244; Verdier, 1990, Yeast 6:271-297).
- the nucleic acid sequence encoding an activator may be obtained from the genes encoding Bacillus stearothermophilus NprA (nprA), Saccharomyces cerevisiae heme activator protein 1 (hapl), Saccharomyces cerevisiae galactose metabolizing protein 4 (gal4), and Aspergillus nidulans ammonia regulation protein (areA).
- nprA Bacillus stearothermophilus NprA
- hapl Saccharomyces cerevisiae heme activator protein 1
- gal4 Saccharomyces cerevisiae galactose metabolizing protein 4
- areA Aspergillus nidulans ammonia regulation protein
- a chaperone is a protein which assists another polypeptide in folding properly (Haiti et al., 1994, TIBS 19:20-25; Bergeron et al, 1994, TIBS 19:124-128; Demolder et al., 1994, Journal of Biotechnology 32:179-189; Craig, 1993, Science 260:1902-1903; Gething and Sambrook, 1992, Nature 355:33-45; Puig and Gilbert, 1994, Journal of Biological Chemistry 269:7764- 5 7771; Wang and Tsou, 1993, The FASEB Journal 7:1515-11157; Robinson et al., 1994, Bio/Technology 1 :381-384).
- the nucleic acid sequence encoding a chaperone may be obtained from the genes encoding Bacillus subtilis GroE proteins, Aspergillus oryzae protein disulphide isomerase, Saccharomyces cerevisiae calnexin, Saccharomyces cerevisiae B ⁇ P/GRP78, and Saccharomyces cerevisiae Hsp70. For further examples, see Gething and o Sambrook, 1992, supra, and Hartl et al., 1994, supra.
- a processing protease is a protease that cleaves a propeptide to generate a mature biochemically active polypeptide (Enderlin and Ogrydziak, 1994, Yeast 10:67-79; Fuller et al., 1989, Proceedings of the National Academy of Sciences USA 86:1434-1438; Julius et al, s 1984, Cell 37:1075-1089; Julius et al., 1983, Cell 32:839-852).
- the nucleic acid sequence encoding a processing protease may be obtained from the genes encoding Aspergillus niger Kex2, Saccharomyces cerevisiae dipeptidylaminopeptidase, Saccharomyces cerevisiae Kex2, and Yanowia lipolytica dibasic processing endoprotease (xpr ⁇ ).
- regulatory sequences which allow the regulation of the expression of the polypeptide relative to the growth of the host cell.
- regulatory systems are those which cause the expression of the gene to be turned on or off in response to a chemical or physical stimulus, including the presence of a regulatory compound.
- Regulatory systems in prokaryotic systems would include the lac, tac, and trp operator systems.
- yeast 5 the ADH2 system or GAL1 system may be used.
- filamentous fungi the TAKA alpha- amylase promoter, Aspergillus niger glucoamylase promoter, and the Aspergillus oryzae glucoamylase promoter may be used as regulatory sequences.
- regulatory sequences are those which allow for gene amplification. In eukaryotic systems, these include the dihydrofolate reductase gene which is amplified in the presence of methotrexate, and the 0 metallothionein genes which are amplified with heavy metals. In these cases, the nucleic acid sequence encoding the polypeptide would be placed in tandem with the regulatory sequence. Promoters
- Suitable promoters for directing the transcription of the nucleic acid constructs of the present invention are the promoters obtained from the E. coli lac operon, the Streptomyces coelicolor agarase gene (dagA), the Bacillus subtilis levansucrase gene (sacB), the Bacillus subtilis alkaline protease gene, the Bacillus licheniformis alpha-amylase gene (amyL), the Bacillus stearothermophilus maltogenic amylase gene (amyM), the Bacillus amyloliquefaciens alpha-amylase gene (amyQ), the Bacillus amyloliquefaciens BAN amylase gene, the Bacillus licheniformis penicillinase gene (penP), the Bacillus subtilis xylA and xylB genes, and the prokaryotic beta-lactamase gene (Villa-Kamaroff
- promoters for directing the transcription of the nucleic acid constructs of the present invention in a filamentous fungal host cell are promoters obtained from the genes encoding Aspergillus oryzae TAKA amylase, Rhizomucor miehei aspartic proteinase, Aspergillus niger neutral alpha-amylase, Aspergillus niger acid stable alpha-amylase, Aspergillus niger or Aspergillus awamori glucoamylase (glaA), Rhizomucor miehei lipase, Aspergillus oryzae alkaline protease, Aspergillus oryzae triose phosphate isomerase, Aspergillus nidulans acetamidase, Fusarium oxysporum trypsin-like protease (as described in U.S.
- Patent No. 4,288,627 which is incorporated herein by reference
- Particularly prefened promoters for use in filamentous fungal host cells are the TAKA amylase, NA2-tpi (a hybrid of the promoters from the genes encoding Aspergillus niger neutral (-amylase and Aspergillus oryzae triose phosphate isomerase), and glaA promoters.
- Further suitable promoters for use in filamentous fungus host cells are the ADH3 promoter (McKnight et al, The EMBO J. 4 (1985), 2093 - 2099) or the tpiA promoter.
- promoters for use in yeast host cells include promoters from yeast glycolytic genes (Hitzeman et al., J. Biol. Chem. 255 (1980), 12073 - 12080; Alber and Kawasaki, J. Mol. Appl. Gen. 1 (1982), 419 - 434) or alcohol dehydrogenase genes (Young et al., in Genetic Engineering of Microorganisms for Chemicals (Hollaender et al, eds.), Plenum Press, New York, 1982), or the TPI1 (US 4,599,311) or ADH2-4c (Russell et al., Nature 304 (1983), 652 - 654) promoters.
- yeast host cells are described by Romanos et al., 1992, Yeast 8:423-488.
- useful promoters include viral promoters such as those from Simian Virus 40 (SV40), Rous sarcoma virus (RSV), adenovirus, and bovine papilloma virus (BPV).
- SV40 Simian Virus 40
- RSV Rous sarcoma virus
- BPV bovine papilloma virus
- Suitable promoters for directing the transcription of the DNA encoding the polypeptide of the invention in mammalian cells are the SV40 promoter (Subramani et al., Mol. Cell Biol. 1 (1981), 854 -864), the MT-1 (metallothionein gene) promoter (Palmiter et al., Science 222 (1983), 809 - 814) or the adenovirus 2 major late promoter.
- An example of a suitable promoter for use in insect cells is the polyhedrin promoter (US 4,745,051; Vasuvedan et al., FEBS Lett. 311, (1992) 7 - 11), the P10 promoter (J.M. Vlak et al., J. Gen.
- Prefened terminators for filamentous fungal host cells are obtained from the genes encoding Aspergillus oryzae TAKA amylase, Aspergillus niger glucoamylase, Aspergillus nidulans anthranilate synthase, Aspergillus niger alpha-glucosidase, and Fusarium oxysporum trypsin- like protease, for fungal hosts) the TPI1 (Alber and Kawasaki, op. cit.) or ADH3 (McKnight et al., op. cit.) terminators.
- TPI1 Alber and Kawasaki, op. cit.
- ADH3 McKnight et al., op. cit.
- Prefened terminators for yeast host cells are obtained from the genes encoding Saccharomyces cerevisiae enolase, Saccharomyces cerevisiae cytochrome C (CYC1), or Saccharomyces cerevisiae glyceraldehyde-3 -phosphate dehydrogenase.
- Other useful terminators for yeast host cells are described by Romanos et al, 1992, supra.
- polyadenylation sequences for filamentous fungal host cells are obtained from the genes encoding Aspergillus oryzae TAKA amylase, Aspergillus niger glucoamylase, Aspergillus nidulans anthranilate synthase, and Aspergillus niger alpha-glucosidase.
- Useful polyadenylation sequences for yeast host cells are described by Guo and Sherman, 1995, Molecular Cellular Biology 15:5983-5990. Polyadenylation sequences are well known in the art for mammalian host cells such as SV40 or the adenovirus 5 Elb region.
- An effective signal peptide coding region for bacterial host cells is the signal peptide coding region obtained from the maltogenic amylase gene from Bacillus NCEB 11837, the Bacillus stearothermophilus alpha-amylase gene, the Bacillus licheniformis subtilisin gene, the Bacillus licheniformis beta-lactamase gene, the Bacillus stearothermophilus neutral proteases genes (nprT, nprS, nprM), and the Bacillus subtilis PrsA gene. Further signal peptides are described by Simonen and Palva, 1993, Microbiological Reviews 57:109-137.
- An effective signal peptide coding region for filamentous fungal host cells is the signal peptide coding region obtained from Aspergillus oryzae TAKA amylase gene, Aspergillus niger neutral amylase gene, the Rhizomucor miehei aspartic proteinase gene, the Humicola lanuginosa cellulase or lipase gene, or the Rhizomucor miehei lipase or protease gene, Aspergillus sp. amylase or glucoamylase, a gene encoding a Rhizomucor miehei lipase or protease.
- the signal peptide is preferably derived from a gene encoding A. oryzae TAKA amylase, A. niger neutral (-amylase, A. niger acid-stable amylase, or A. niger glucoamylase.
- Useful signal peptides for yeast host cells are obtained from the genes for Saccharomyces cerevisiae a- factor and Saccharomyces cerevisiae invertase. Other useful signal peptide coding regions are described by Romanos et al., 1992, supra.
- the secretory signal sequence may encode any signal peptide which ensures efficient direction of the expressed polypeptide into the secretory pathway of the cell.
- the signal peptide may be naturally occurring signal peptide, or a functional part thereof, or it may be a synthetic peptide. Suitable signal peptides have been found to be the a- factor signal peptide (cf.
- a sequence encoding a leader peptide may also be inserted downstream of the signal sequence and uptream of the DNA sequence encoding the polypeptide.
- the function of the leader peptide is to allow the expressed polypeptide to be directed from the endoplasmic reticulum to the Golgi apparatus and further to a secretory vesicle for secretion into the culture medium (i.e. exportation of the polypeptide across the cell wall or at least through the cellular membrane into the periplasmic space of the yeast cell).
- the leader peptide may be the yeast a- factor leader (the use of which is described in e.g.
- the leader peptide may be a synthetic leader peptide, which is to say a leader peptide not found in nature. Synthetic leader peptides may, for instance, be constructed as described in WO 89/02463 or WO 92/11378.
- the signal peptide may conveniently be derived from an insect gene (cf. WO 90/05783), such as the lepidopteran Manduca sexta adipokinetic hormone precursor signal peptide (cf. US 5,023,328).
- the present invention also relates to recombinant expression vectors comprising a nucleic acid sequence of the present invention, a promoter, and transcriptional and translational stop signals.
- the various nucleic acid and control sequences described above may be joined together to produce a recombinant expression vector which may include one or more convenient restriction sites to allow for insertion or substitution of the nucleic acid sequence encoding the polypeptide at such sites.
- the nucleic acid sequence of the present invention may be expressed by inserting the nucleic acid sequence or a nucleic acid construct comprising the sequence into an appropriate vector for expression.
- the coding sequence is located in the vector so that the coding sequence is operably linked with the appropriate control sequences for expression, and possibly secretion.
- the recombinant expression vector may be any vector (e.g., a plasmid or virus) which can be conveniently subjected to recombinant DNA procedures and can bring about the expression of the nucleic acid sequence.
- the choice of the vector will typically depend on the compatibility of the vector with the host cell into which the vector is to be introduced.
- the vectors may be linear or closed circular plasmids.
- the vector may be an autonomously replicating vector, i.e., a vector which exists as an extrachromosomal entity, the replication of which is independent of chromosomal replication, e.g., a plasmid, an extrachromosomal element, a minichromosome, or an artificial chromosome.
- the vector may contain any means for assuring self-replication.
- the vector may be one which, when introduced into the host cell, is integrated into the genome and replicated together with the chromosome(s) into which it has been integrated.
- the vector system may be a single vector or plasmid or two or more vectors or plasmids which together contain the total DNA to be introduced into the genome of the host cell, or a transposon.
- the vectors of the present invention preferably contain one or more selectable markers which permit easy selection of transformed cells.
- a selectable marker is a gene the product of which provides for biocide or viral resistance, resistance to heavy metals, prototrophy to auxotrophs, and the like.
- Examples of bacterial selectable markers are the dal genes from Bacillus subtilis or Bacillus licheniformis, or markers which confer antibiotic resistance such as ampicillin, kanamycin, chloramphenicol, tetracycline, neomycin, hygromycin or methotrexate resistance.
- a frequently used mammalian marker is the dihydrofolate reductase gene (DHFR).
- Suitable markers for yeast host cells are ADE2, HIS3, LEU2, LYS2, MET3, TRP1, and URA3.
- a selectable marker for use in a filamentous fungal host cell may be selected from the group including, but not limited to, amdS (acetamidase), argB (ornithine carbamoyltransferase), bar (phosphinothricin acetyltransferase), hygB (hygromycin phosphotransferase), niaD (nitrate reductase), pyrG (orotidine-5' -phosphate decarboxylase), sC (sulfate adenyltransferase), trpC (anthranilate synthase), and glufosinate resistance markers, as well as equivalents from other species.
- amdS and pyrG markers of Aspergillus nidulans or Aspergillus oryzae are the amdS and pyrG markers of Aspergillus nidulans or Aspergillus oryzae and the bar marker of Streptomyces hygroscopicus. Furthermore, selection may be accomplished by co-transformation, e.g., as described in WO 91/17243, where the selectable marker is on a separate vector.
- the vectors of the present invention preferably contain an element(s) that permits stable integration of the vector into the host cell genome or autonomous replication of the vector in the cell independent of the genome of the cell.
- the vectors of the present invention may be integrated into the host cell genome when introduced into a host cell.
- the vector may rely on the nucleic acid sequence encoding the polypeptide or any other element of the vector for stable integration of the vector into the genome by homologous or nonhomologous recombination.
- the vector may contain additional nucleic acid sequences for directing integration by homologous recombination into the genome of the host cell. The additional nucleic acid sequences enable the vector to be integrated into the host cell genome at a precise location(s) in the chromosome(s).
- the integrational elements should preferably contain a sufficient number of nucleic acids, such as 100 to 1,500 base pairs, preferably 400 to 1,500 base pairs, and most preferably 800 to 1,500 base pairs, which are highly homologous with the conesponding target sequence to enhance the probability of homologous recombination.
- the integrational elements may be any sequence that is homologous with the target sequence in the genome of the host cell.
- the integrational elements may be non-encoding or encoding nucleic acid sequences.
- the vector may be integrated into the genome of the host cell by non-homologous recombination.
- These nucleic acid sequences may be any sequence that is homologous with a target sequence in the genome of the host cell, and, furthermore, may be non-encoding or encoding sequences.
- the vector may further comprise an origin of replication enabling the vector to replicate autonomously in the host cell in question.
- origins of replication are the origins of replication of plasmids pBR322, pUC19, pACYC177, pACYC184, pUBHO, pE194, pTA1060, and pAMBl.
- origin of replications for use in a yeast host cell are the 2 micron origin of replication, the combination of CEN6 and ARS4, and the combination of CEN3 and ARS1.
- the origin of replication may be one having a mutation which makes its functioning temperature-sensitive in the host cell (see, e.g., Ehrlich, 1978, Proceedings of the National Academy of Sciences USA 75:1433).
- More than one copy of a nucleic acid sequence encoding a polypeptide of the present invention may be inserted into the host cell to amplify expression of the nucleic acid sequence.
- Stable amplification of the nucleic acid sequence can be obtained by integrating at least one additional copy of the sequence into the host cell genome using methods well known in the art and selecting for transformants.
- the present invention also relates to recombinant host cells, comprising a nucleic acid sequence of the invention, which are advantageously used in the recombinant production of the polypeptides.
- host cell encompasses any progeny of a parent cell which is not identical to the parent cell due to mutations that occur during replication.
- the cell is preferably transformed with a vector comprising a nucleic acid sequence of the invention followed by integration of the vector into the host chromosome.
- Transformation means introducing a vector comprising a nucleic acid sequence of the present invention into a host cell so that the vector is maintained as a chromosomal integrant or as a self-replicating extra-chromosomal vector. Integration is generally considered to be an advantage as the nucleic acid sequence is more likely to be stably maintained in the cell. Integration of the vector into the host chromosome may occur by homologous or non-homologous recombination as described above.
- the choice of a host cell will to a large extent depend upon the gene encoding the polypeptide and its source.
- the host cell may be a unicellular microorganism, e.g., a prokaryote, or a non- unicellular microorganism, e.g., a eukaryote.
- Useful unicellular cells are bacterial cells such as gram positive bacteria including, but not limited to, a Bacillus cell, e.g., Bacillus alkalophilus, Bacillus amyloliquefaciens, Bacillus brevis, Bacillus circulans, Bacillus coagulans, Bacillus lautus, Bacillus lentus, Bacillus licheniformis, Bacillus megaterium, Bacillus stearothermophilus, Bacillus subtilis, and Bacillus thuringiensis; or a Streptomyces cell, e.g., Streptomyces lividans or Streptomyces murinus, or gram negative bacteria such as E.
- a Bacillus cell e.g., Bacillus alkalophilus, Bacillus amyloliquefaciens, Bacillus brevis, Bacillus circulans, Bacillus coagulans, Bacillus lautus, Bacillus lentus, Bacillus licheniformis
- the bacterial host cell is a Bacillus lentus, Bacillus licheniformis, Bacillus stearothermophilus or Bacillus subtilis cell.
- the transformation of a bacterial host cell may, for instance, be effected by protoplast transformation (see, e.g., Chang and Cohen, 1979, Molecular General Genetics 168:111-115), by using competent cells (see, e.g., Young and Spizizin, 1961, Journal of Bacteriology 81 :823-829, or Dubnar and Davidoff-Abelson, 1971, Journal of Molecular Biology 56:209- 221), by electroporation (see, e.g., Shigekawa and Dower, 1988, Biotechniques 6:742-751), or by conjugation (see, e.g., Koehler and Thome, 1987, Journal of Bacteriology 169:5771-5278).
- protoplast transformation see, e.g., Chang and Cohen, 1979, Molecular General Genetics 168:111-115
- competent cells see, e.g., Young and Spizizin, 1961, Journal of Bacteriology 81 :823-829, or Dubnar and
- the host cell may be a eukaryote, such as a mammalian cell, an insect cell, a plant cell or a fungal cell.
- Useful mammalian cells include Chinese hamster ovary (CHO) cells, HeLa cells, baby hamster kidney (BHK) cells, COS cells, or any number of other immortalized cell lines available, e.g., from the American Type Culture Collection.
- suitable mammalian cell lines are the COS (ATCC CRL 1650 and 1651), BHK (ATCC CRL 1632, 10314 and 1573, ATCC CCL 10), CHL (ATCC CCL39) or CHO (ATCC CCL 61) cell lines.
- Methods of transfecting mammalian cells and expressing DNA sequences introduced in the cells are described in e.g. Kaufman and Sharp, J. Mol. Biol. 159 (1982), 601 - 621; Southern and Berg, J. Mol. Appl. Genet. 1 (1982), 327 - 341; Loyter et al, Proc. Natl. Acad. Sci.
- the host cell is a fungal cell.
- Fungi as used herein includes the phyla Ascomycota, Basidiomycota, Chytridiomycota, and Zygomycota (as defined by Hawksworth et al., In, Ainsworth and Bisby's Dictionary of The Fungi, 8th edition, 1995, CAB International, University Press, Cambridge, UK) as well as the Oomycota (as cited in Hawksworth et al., 1995, supra, page 171) and all mitosporic fungi (Hawksworth et al., 1995, supra).
- Examples of Basidiomycota include mushrooms, rusts, and smuts.
- Representative groups of Chytridiomycota include, e.g., Allomyces, Blastocladiella, Coelomomyces, and aquatic fungi.
- Representative groups of Oomycota include, e.g., Saprolegniomycetous aquatic fungi (water molds) such as Achlya.
- mitosporic fungi examples include Aspergillus, Penicillium, Candida, and Alternaria.
- Representative groups of Zygomycota include, e.g., Rhizopus and Mucor.
- the fungal host cell is a yeast cell.
- yeast as used herein includes ascosporogenous yeast (Endomycetales), basidiosporogenous yeast, and yeast belonging to the Fungi Imperfecti (Blastomycetes). The ascosporogenous yeasts are divided into the families Spermophthoraceae and Saccharomycetaceae.
- the latter is comprised of four subfamilies, Schizosaccharomycoideae (e.g., genus Schizosaccharomyces), Nadsonioideae, Lipomycoideae, and Saccharomycoideae (e.g., genera Pichia, Kluyveromyces and Saccharomyces).
- the basidiosporogenous yeasts include the genera Leucosporidim, Rhodosporidium, Sporidiobolus, Filobasidium, and Filobasidiella.
- yeast belonging to the Fungi Imperfecti are divided into two families, Sporobolomycetaceae (e.g., genera Sorobolomyces and Bullera) and Cryptococcaceae (e.g., genus Candida). Since the classification of yeast may change in the future, for the purposes of this invention, yeast shall be defined as described in Biology and Activities of Yeast (Skinner, F.A., Passmore, S.M., and Davenport, R.R., eds, Soc. App. Bacteriol. Symposium Series No. 9, 1980.
- yeast host cell may be selected from a cell of a species of Candida, Kluyveromyces, Saccharomyces, Schizosaccharomyces, Candida, Pichia, Hansehula, , or Yanowia.
- the yeast host cell is a Saccharomyces carlsbergensis, Saccharomyces cerevisiae, Saccharomyces diastaticus, Saccharomyces douglasii, Saccharomyces kluyveri, Saccharomyces norbensis or Saccharomyces oviformis cell.
- Other useful yeast host cells are a Kluyveromyces lactis Kluyveromyces fragilis Hansehula polymorpha, Pichia pastoris Yanowia lipolytica, Schizosaccharomyces pombe, Ustilgo maylis, Candida maltose, Pichia guillermondii and Pichia methanolio cell (cf. Gleeson et al., J. Gen. Microbiol. 132, 1986, pp. 3459-3465; US 4,882,279 and US 4,879,231).
- the fungal host cell is a filamentous fungal cell.
- filamentous fungi include all filamentous forms of the subdivision Eumycota and Oomycota (as defined by Hawksworth et al, 1995, supra).
- the filamentous fungi are characterized by a vegetative mycelium composed of chitin, cellulose, glucan, chitosan, mannan, and other complex polysaccharides.
- Vegetative growth is by hyphal elongation and carbon catabolism is obligately aerobic.
- vegetative growth by yeasts such as Saccharomyces cerevisiae is by budding of a unicellular thallus and carbon catabolism may be fermentative.
- the filamentous fungal host cell is a cell of a species of, but not limited to, Acremonium, Aspergillus, Fusarium, Humicola, Mucor, Myceliophthora, Neurospora, Penicillium, Thielavia, Tolypocladium, and Trichoderma or a teleomorph or synonym thereof.
- the filamentous fungal host cell is an Aspergillus cell.
- the filamentous fungal host cell is an Acremonium cell.
- the filamentous fungal host cell is a Fusarium cell.
- the filamentous fungal host cell is a Humicola cell. In another even more prefened embodiment, the filamentous fungal host cell is a Mucor cell. In another even more prefened embodiment, the filamentous fungal host cell is a Myceliophthora cell. In another even more prefened embodiment, the filamentous fungal host cell is a Neurospora cell. In another even more prefened embodiment, the filamentous fungal host cell is a Penicillium cell. In another even more prefened embodiment, the filamentous fungal host cell is a Thielavia cell. In another even more prefened embodiment, the filamentous fungal host cell is a Tolypocladium cell.
- the filamentous fungal host cell is a Trichoderma cell.
- the filamentous fungal host cell is an Aspergillus awamori, Aspergillus foetidus, Aspergillus japonicus, Aspergillus niger, Aspergillus nidulans or Aspergillus oryzae cell.
- the filamentous fungal host cell is a Fusarium cell of the section Discolor (also known as the section Fusarium).
- the filamentous fungal parent cell may be a Fusarium bactridioides, Fusarium cerealis, Fusarium crookwellense, Fusarium culmorum, Fusarium graminearum, Fusarium graminum, Fusarium heterosporum, Fusarium negundi, Fusarium reticulatum, Fusarium roseum, Fusarium sambucinum, Fusarium sarcochroum, Fusarium sulphureum, or Fusarium trichothecioides cell.
- Fusarium bactridioides Fusarium cerealis, Fusarium crookwellense, Fusarium culmorum, Fusarium graminearum, Fusarium graminum, Fusarium heterosporum, Fusarium negundi, Fusarium reticulatum, Fusarium roseum, Fusarium sambucinum, Fusarium sarc
- the filamentous fungal parent cell is a Fusarium strain of the section Elegans, e.g., Fusarium oxysporum.
- the filamentous fungal host cell is a Humicola insolens or Humicola lanuginosa cell.
- the filamentous fungal host cell is a Mucor miehei cell.
- the filamentous fungal host cell is a Myceliophthora thermophilum cell.
- the filamentous fungal host cell is a Neurospora crassa cell.
- the filamentous fungal host cell is a Penicillium purpurogenum cell.
- the filamentous fungal host cell is a Thielavia tenestris cell or a Acremonium chrysogenum cell.
- the Trichoderma cell is a Trichoderma harzianum, Trichoderma koningii, Trichoderma longibrachiatum, Trichoderma reesei or Trichoderma viride cell.
- Aspergillus spp. for the expression of proteins is described in, e.g., EP 272 277, EP 230 023.
- Fungal cells may be transformed by a process involving protoplast formation, transformation of the protoplasts, and regeneration of the cell wall in a manner known per se. Suitable procedures for transformation of Aspergillus host cells are described in EP 238 023 and Yelton et al, 1984, Proceedings of the National Academy of Sciences USA 81:1470-1474. A suitable method of transforming Fusarium species is described by Malardier et al., 1989, Gene 78:147-156 or in copending US Serial No. 08/269,449. Examples of other fungal cells are cells of filamentous fungi, e.g. Aspergillus spp., Neurospora spp., Fusarium spp. or
- Trichoderma spp. in particular strains of A. oryzae, A. nidulans or A. niger.
- the use of Aspergillus spp. for the expression of proteins is described in, e.g., EP 272 277, EP 230 023, EP 184 ...
- the transformation of F. oxysporum may, for instance, be carried out as described by Malardier et al., 1989, Gene 78: 147-156.
- Yeast may be transformed using the procedures described by Becker and Guarente, In Abelson, J.N. and Simon, M.I., editors, Guide to Yeast Genetics and Molecular Biology, Methods in Enzymology, Volume 194, pp 182-187, Academic Press, Inc., New York; Ito et al., 1983, Journal of Bacteriology 153:163; and Hinnen et al., 1978, Proceedings of the National Academy of Sciences USA 75:1920. Mammalian cells may be transformed by direct uptake using the calcium phosphate precipitation method of Graham and Van der Eb (1978, Virology 52:546).
- Transformation of insect cells and production of heterologous polypeptides therein may be performed as described in US 4,745,051; US 4, 775, 624; US 4,879,236; US 5,155,037; US 5,162,222; EP 397,485) all of which are inco ⁇ orated herein by reference.
- the insect cell line used as the host may suitably be a Lepidoptera cell line, such as Spodoptera frugiperda cells or Trichoplusia ni cells (cf. US 5,077,214).
- Culture conditions may suitably be as described in, for instance, WO 89/01029 or WO 89/01028, or any of the aforementioned references.
- the transformed or transfected host cells described above are cultured in a suitable nutrient medium under conditions permitting the production of the desired molecules, after which these are recovered from the cells, or the culture broth.
- the medium used to culture the cells may be any conventional medium suitable for growing the host cells, such as minimal or complex media containing appropriate supplements. Suitable media are available from commercial suppliers or may be prepared according to published recipes (e.g. in catalogues of the American Type Culture Collection). The media are prepared using procedures known in the art (see, e.g., references for bacteria and yeast; Bennett, J.W. and LaSure, L., editors, More Gene Manipulations in Fungi, Academic Press, CA, 1991). If the molecules are secreted into the nutrient medium, they can be recovered directly from the medium. If they are not secreted, they can be recovered from cell lysates.
- the molecules are recovered from the culture medium by conventional procedures including separating the host cells from the medium by centrifugation or filtration, precipitating the proteinaceous components of the supernatant or filtrate by means of a salt, e.g. ammonium sulphate, purification by a variety of chromatographic procedures, e.g. ion exchange chromatography, gelfiltration chromatography, affinity chromatography, or the like, dependent on the type of molecule in question.
- a salt e.g. ammonium sulphate
- the molecules of interest may be detected using methods known in the art that are specific for the molecules. These detection methods may include use of specific antibodies, formation of a product, or disappearance of a substrate. For example, an enzyme assay may be used to determine the activity of the molecule. Procedures for determining various kinds of activity are known in the art.
- the molecules of the present invention may be purified by a variety of procedures known in the art including, but not limited to, chromatography (e.g., ion exchange, affinity, hydrophobic, chromatofocusing, and size exclusion), electrophoretic procedures (e.g., preparative isoelectric focusing (IEF), differential solubility (e.g., ammonium sulfate precipitation), or extraction (see, e.g., Protein Purification, J-C Janson and Lars Ryden, editors, VCH Publishers, New York, 1989).
- chromatography e.g., ion exchange, affinity, hydrophobic, chromatofocusing, and size exclusion
- electrophoretic procedures e.g., preparative isoelectric focusing (IEF), differential solubility (e.g., ammonium sulfate precipitation), or extraction
- IEF isoelectric focusing
- differential solubility e.g., ammonium sulfate precipitation
- extraction see, e.g.
- immunological response is the response of an organism to a compound, which involves the immune system according to any of the four standard reactions (Type I, II, III and IV according to Coombs & Gell).
- the "immunogenicity" of a compound used in connection with the present invention refers to the ability of this compound to induce an 'immunological response' in animals including man.
- allergic response is the response of an organism to a compound, which involves IgE mediated responses (Type I reaction according to Coombs & Gell). It is to be understood that sensibilization (i.e. development of compound-specific IgE antibodies) upon exposure to the compound is included in the definition of "allergic response”.
- the "allergenicity" of a compound used in connection with the present 5 invention refers to the ability of this compound to induce an 'allergic response' in animals including man.
- relevant protein backbone or 'protein backbone' refer to the polypeptide to be modified by creating a library of diversified mutants.
- the "relevant protein backbone” may be o a naturally occurring (or wild-type) polypeptide or it may be a variant thereof prepared by any suitable means.
- the “relevant protein backbone” may be a variant of a naturally occurring polypeptide which has been modified by substitution, deletion or truncation of one or more amino acid residues or by addition or insertion of one or more amino acid residues to the amino acid sequence of a naturally-occurring polypeptide. 5
- randomized library of protein variants refers to a library with at least partially randomized composition of the members, e.g. protein variants.
- the term "functionality" of protein variants refers to e.g. enzymatic activity, binding to a 0 ligand or receptor, stimulation of a cellular response (e.g. 3 H-thymidine incorporation as response to a mitogenic factor), or anti-microbial activity.
- an “epitope” is a set of amino acids on a protein that are involved in an immunological response, such as antibody binding or T-cell activation.
- One particularly useful method of 5 identifying epitopes involved in antibody binding is to screen a library of peptide-phage membrane protein fusions and selecting those that bind to relevant antigen-specific antibodies, sequencing the randomized part of the fusion gene, aligning the sequences involved in binding, defining consensus sequences based on these alignments, and mapping these consensus sequences on the surface or the sequence and/or structure of the antigen, to identify o epitopes involved in antibody binding.
- epitopes involved in antibody binding is a consensus sequence of peptides that bind well to a relevant antibody.
- epitope area is defined as the amino acids situated within 5 A from the epitope amino acids. Modifications of amino acids of the 'epitope area' can possibly affect the function of the conesponding epitope.
- polyclonal antibodies polyclonal antibodies isolated according to their specificity for a certain antigen, e.g. the protein backbone.
- monospecific antibodies polyclonal antibodies isolated according to their specificity for a certain 'epitope pattern'. Such monospecific antibodies will only bind to one epitope pattern, but they may very well be produced by a number of antibody producing cells and recognize the same epitope patterns, thereby being polyclonal.
- 'Spiked mutagenesis' is a form of site-directed mutagenesis, in which the primers used have been synthesized using mixtures of oligonucleotides at one or more positions.
- the inventors have found a method for high throughput screening (HTS) of a large population of host cells for production of a molecule of interest.
- HTS high throughput screening
- the present invention relates to method for screening a library of protein variants for functional variants with reduced antibody binding capacity, comprising the steps of: (i) generating a diversified library of protein variants starting from a relevant protein backbone,
- the "relevant protein backbone” can in principle be any protein molecule of biological origin, non-limiting examples of which are peptides, polypeptides, proteins, enzymes, post- translationally modified polypeptides such as lipopeptides or glycosylated peptides, antimicrobial peptides or molecules, and proteins having pharmaceutical properties etc.
- the invention relates to a method, wherein the "relevant protein backbone" is chosen from the group consisting of polypeptides, small peptides, lipopeptides, antimicrobials, and pharmaceutical polypeptides.
- pharmaceutical polypeptides is defined as polypeptides, including peptides, such as peptide hormones, proteins and/or enzymes, being physiologically active when administered to humans and/or animals.
- Examples of “pharmaceutical polypeptides” contemplated according to the invention include insulin, ACTH, glucagon, somatostatin, somatotropin, thymosin, parathyroid hormone, pigmentary hormones, somatomedin, erythropoietin, luteinizing hormone, chorionic gonadotropin, hypothalmic releasing factors, antidiuretic hormones, thyroid stimulating hormone, relaxin, interferons, thrombopoietin (TPO), blood coagulation factors, plasminogen activators such as streptokinase and tissue-plasminogen activator, cerebrosidase, and prolactin.
- proteins are preferably to be used in industry, housekeeping and/or medicine, such as proteins used in personal care products (for example shampoo; soap; skin, hand and face lotions; skin, hand and face creams; hair dyes; toothpaste), food/feed (for example in the baking industry), detergents, anti-microbial compositions.
- personal care products for example shampoo; soap; skin, hand and face lotions; skin, hand and face creams; hair dyes; toothpaste
- food/feed for example in the baking industry
- detergents for example in the baking industry
- the protein is an enzyme, such as glycosyl hydrolases, carbohydrases, peroxidases, proteases, lipases, phytases, polysaccharide lyases, oxidoreductases, transglutaminases and glycoseisomerases, in particular the following.
- enzymes such as glycosyl hydrolases, carbohydrases, peroxidases, proteases, lipases, phytases, polysaccharide lyases, oxidoreductases, transglutaminases and glycoseisomerases, in particular the following.
- IUBMB Molecular Biology
- proteases selected from those classified under the Enzyme Classification
- 3.4.11 i.e. so-called aminopeptidases
- aminopeptidases including 3.4.11.5 (Prolyl aminopeptidase), 3.4.11.9 (X-pro aminopeptidase), 3.4.11.10 (Bacterial leucyl aminopeptidase), 3.4.11.12 (Thermophihc aminopeptidase), 3.4.11.15 (Lysyl aminopeptidase), 3.4.11.17 (Tryptophanyl aminopeptidase),
- 3.4.21 i.e. so-called serine endopeptidases
- 3.4.21.1 Chomotrypsin
- 3.4.21.4 Trypsin
- 3.4.21.25 Cucumisin
- 3.4.21.32 Brachyurin
- 3.4.21.48 Cerevisin
- 3.4.21.62 Subtilisin
- 3.4.22 i.e. so-called cysteine endopeptidases
- cysteine endopeptidases including 3.4.22.2 (Papain), 3.4.22.3 (Ficain), 3.4.22.6 (Chymopapain), 3.4.22.7 (Asclepain), 3.4.22.14 (Actinidain), 3.4.22.30 (Caricain) and 3.4.22.31 (Ananain);
- 3.4.23 i.e. so-called aspartic endopeptidases
- aspartic endopeptidases including 3.4.23.1 (Pepsin A), 3.4.23.18 (Aspergillopepsin I), 3.4.23.20 (Penicillopepsin) and 3.4.23.25 (Saccharopepsin); and
- subtilisins comprise subtilisin BPN', subtilisin amylosacchariticus, subtilisin 168, subtilisin mesentericopeptidase, subtilisin Carlsberg, subtilisin DY, subtilisin 309, subtilisin 147, the ⁇ nitase, aqualysin, Bacillus PB92 protease, proteinase K, Protease TW7, and Protease TW3.
- proteases include Esperase®, Alcalase®, Neutrase®, Dyrazym®, Savinase®, Pyrase®, Pancreatic Trypsin NOVO (PTN), Bio-Feed( Pro, Clear-Lens Pro (all enzymes available from Novo Nordisk A/S).
- examples of other commercial proteases include Maxtase®, Maxacal®, Maxapem® marketed by Gist-Brocades N.V., Opticlean® marketed by Solvay et Cie. and Purafect® marketed by l o Genencor International .
- protease variants are contemplated as the parent protease.
- protease variants are disclosed in EP 130.756 (Genentech), EP 214.435 (Henkel), WO 87/04461 (Amgen), WO 87/05050 (Genex), EP 251.446 (Genencor), EP 260.105 (Genencor), Thomas et al., (1985), Nature. 318, p. 375-376, Thomas et al., (1987), J. is Mol. Biol., 193, pp. 803-813, Russel et al., (1987), Nature, 328, p.
- Parent lipases i.e. enzymes classified under the Enzyme Classification number E.C. 3.1.1
- IUBMB International Union of Biochemistry and Molecular Biology
- Examples include lipases selected from those classified under the Enzyme Classification
- lipases include lipases derived from the following microorganisms.
- Humicola e.g. H. brevispora, H. lanuginosa, H. brevis var. thermoidea and H. insolens (US)
- Pseudomonas e.g. Ps. fragi, Ps. stutzeri, Ps. cepacia and Ps. fluorescens (WO 89/04361), or
- Ps. plantarii or Ps. gladioli US patent no. 4,950,417 (Solvay enzymes)
- Ps. alcaligenes and 5 Ps. pseudoalcaligenes EP 218 272
- Ps. mendocina WO 88/09367; US 5,389,536).
- Fusarium e.g. F. oxyspomm (EP 130,064) or F. solani pisi(WO 90/09446).
- Mucor also called Rhizomucor
- Chromobacterium e.g. M. miehei (EP 238 023).
- C. viscosum especially C. viscosum
- Aspergillus especially A. niger
- Candida e.g. C. cylindracea (also called C. rugosa) or C. antarctica (WO 88/02775) or C. antarctica lipase A or B (WO o 94/01541 and WO 89/02916).
- Geotricum e.g. G. candidum (Schimada et al., (1989), J.
- Penicillium e.g. P. camembertii (Yamaguchi et al., (1991), Gene
- Rhizopus e.g. R. delemar (Hass et al., (1991), Gene 109, 107-113) or R. niveus (Kugimiya et al., (1992) Biosci. Biotech. Biochem 56, 716-719) or R. oryzae.
- Bacillus e.g. B. subtilis (Dartois et al, (1993) s Biochemica et Biophysica acta 1131, 253-260) or B. stearothermophilus (JP 64/7744992) or
- lipases include Lipolase®, Lipolase( Ultra, Lipozyme®, Palatase®, Novozym® 435, Lecitase® (all available from Novo Nordisk A/S).
- lipases are Lumafast(, Ps. mendocian lipase from Genencor Int. Inc.; Lipomax(, Ps. pseudoalcaligenes lipase from Gist Brocades/Genencor Int. Inc.; Fusarium solani lipase (cutinase) from Unilever; Bacillus sp. lipase from Solvay enzymes.
- Other lipases are available from other companies. It is to be understood that also lipase variants are contemplated as the parent enzyme. 5 Examples of such are described in e.g. WO 93/01285 and WO 95/22615.
- the activity of the lipase can be determined as described in "Methods of Enzymatic Analysis", Third Edition, 1984, Verlag Chemie, Weinhein, vol. 4, or as described in AF 95/5 GB (available on request from Novo Nordisk A/S).
- oxidoreductases i.e. enzymes classified under the Enzyme Classification number E.C. 1 (Oxidoreductases) in accordance with the Recommendations (1992) of the International Union of Biochemistry and Molecular Biology (IUBMB)
- E.C. 1 Enzyme Classification number
- IUBMB International Union of Biochemistry and Molecular Biology
- Examples include oxidoreductases selected from those classified under the Enzyme Classification (E.C.) numbers: Glycerol-3-phosphate dehydrogenase _NAD+_ (1.1.1.8), Glycerol-3 -phosphate dehydrogenase _NAD(P)+_ (1.1.1.94), Glycerol-3 -phosphate 1 -dehydrogenase _NADP_ (1.1.1.94), Glucose oxidase (1.1.3.4), Hexose oxidase (1.1.3.5), Catechol oxidase (1.1.3.14), Bilirubin oxidase (1.3.3.5), Alanine dehydrogenase (1.4.1.1), Glutamate dehydrogenase (1.4.1.2), Glutamate dehydrogenase _NAD(P)+_ (1.4.1.3), Glutamate dehydrogenase _NADP+_ (1.4.1.4), L-Amino acid dehydrogenase (1.4
- Said Glucose oxidases may be derived from Aspergillus niger.
- Said Laccases may be derived from Polyporus pinsitus, Myceliophtora thermophila, Coprinus cinereus, Rhizoctonia solani,
- Bilimbin oxidases may be derived from Myrothechecium verrucaria.
- the Peroxidase may be derived from e.g.
- Protein Disulfide reductase may be any of the mentioned in DK patent applications No. 768/93, 265/94 and 264/94 (Novo Nordisk A S), which are hereby incorporated as references, including Protein Disulfide reductases of bovine origin, Protein Disulfide reductases derived from Aspergillus oryzae or Aspergillus niger, and
- DsbA or DsbC derived from Escherichia coli.
- oxidoreductases include Gluzyme (enzyme available from Novo Nordisk A/S). However, other oxidoreductases are available from others. It is to be understood that also variants of oxidoreductases are contemplated as the parent enzyme. The activity of oxidoreductases can be determined as described in "Methods of Enzymatic Analysis", third edition, 1984, Verlag Chemie, Weinheim, vol. 3.
- Parent carbohydrases may be defined as all enzymes capable of breaking down carbohydrate chains (e.g. starches) of especially five and six member ring structures (i.e. enzymes classified under the Enzyme Classification number E.C. 3.2 (glycosidases) in accordance with the Recommendations (1992) of the International Union of Biochemistry and Molecular Biology (IUBMB)).
- E.C. 3.2 glycosidases
- carbohydrases examples include (-1,3-glucanases derived from Trichoderma harzianum; (-1,6-glucanases derived from a strain of Paecilomyces; (-glucanases derived from Bacillus subtilis; (-glucanases derived from Humicola insolens; (-glucan-ases derived from Aspergillus niger; (-glucanases derived from a strain of Trichoderma; (-glucanases derived 5 from a strain of Oerskovia xanthineolytica; exo-l,4-(-D-glucosidases (glucoamylases) derived from Aspergillus niger; (-amylases derived from Bacillus subtilis; (-amylases derived from Bacillus amyloliquefaciens; (-amylases derived from Bacillus stearothermophilus; (-amylases derived from As
- carbohydrases include Alpha-Gal(, Bio- 5 Feed( Alpha, Bio-Feed( Beta, Bio-Feed( Plus, Bio-Feed( Plus, Novozyme® 188, Carezyme®, Celluclast®, Cellusoft®, Ceremyl®, Citrozym(, Denimax(, Dezyme(, Dextrozyme(, Finizym®, Fungamyl(, Gamanase(, Glucanex®, Lactozym®, Maltogenase(, Pentopan(, Pectinex(, Promozyme®, Pulpzyme(, Novamyl(, Termamyl®, AMG (Amyloglucosidase Novo), Maltogenase®, Aquazym®, Natalase( (all enzymes available from Novo Nordisk o A S).
- Other carbohydrases are available from other companies.
- carbohydrase variants are contemplated as the parent enzyme.
- the activity of carbohydrases can be determined as described in "Methods of Enzymatic Analysis", third edition, 1984, Verlag Chemie, Weinheim, vol. 4.
- Parent transferases i.e. enzymes classified under the Enzyme Classification number E.C. 2 in accordance with the Recommendations (1992) of the International Union of Biochemistry and Molecular Biology (IUBMB)
- the parent transferases may be any transferase in the subgroups of transferases: transferases 0 transferring one-carbon groups (E.C. 2.1); transferases transferring aldehyde or residues (E.C 2.2); acyltransferases (E.C. 2.3); glucosyltransferases (E.C. 2.4); transferases transferring alkyl or aryl groups, other that methyl groups (E.C.
- the parent transferase is a transglutaminase E.C 2.3.2.13 (Protein- 5 glutamine (-glutamyltransferase) .
- Transglutaminases are enzymes capable of catalyzing an acyl transfer reaction in which a gamma-carboxyamide group of a peptide-bound glutamine residue is the acyl donor.
- Primary amino groups in a variety of compounds may function as acyl acceptors with the subsequent formation of monosubstituted gamma-amides of peptide-bound glutamic acid.
- the transferases form intramolecular or intermolecular gamma-glutamyl-epsilon-lysyl crosslinks. Examples of transglutaminases are described in the pending DK patent application no. 990/94 (Novo Nordisk A/S).
- the parent transglutaminase may be of human, animal (e.g. bovine) or microbial origin.
- parent transglutaminases are animal derived Transglutaminase, FXUIa; microbial transglutaminases derived from Physarum polycephalum (Klein et al., Journal of Bacteriology, Vol. 174, p.
- transglutaminases derived from Streptomyces sp. including Streptomyces lavendulae, Streptomyces lydicus (former Streptomyces libani) and Streptoverticillium sp., including Streptoverticillium mobaraense, Streptoverticillium cin- namoneum, and Streptoverticillium griseocarneum (Motoki et al., US 5,156,956; Andou et al., US 5,252,469; Kaempfer et al., Journal of General Microbiology, Vol. 137, p.
- transglutaminases can be determined as described in "Methods of Enzymatic Analysis", third edition, 1984, Verlag Chemie, Weinheim, vol. 1-10.
- Phytases are enzymes produced by microorganisms, which catalyse the conversion of phytate to inositol and inorganic phosphorus.
- Phytase producing microorganisms comprise bacteria such as Bacillus subtilis, Bacillus natto and Pseudomonas; yeasts such as Saccharomyces cerevisiae; and fungi such as Aspergillus niger, Aspergillus ficuum, Aspergillus awamori, Aspergillus oryzae, Aspergillus teneus or
- parent phytases examples include phytases selected from those classified under the Enzyme
- Suitable lyases include Polysaccharide lyases: Pectate lyases (4.2.2.2) and pectin lyases
- suitable protein disulfide isomerases include PDIs described in WO 95/01425 (Novo Nordisk A/S) and suitable glucose isomerases include those described in Biotechnology Letter, Vol. 20, No 6, June 1998, pp. 553-56.
- Contemplated isomerases include xylose/glucose Isomerase (5.3.1.5) including Sweetzyme®.
- the methods of this invention are especially suitable when testing compounds that are being modified with respect to their allergenicity.
- Such modification of a test compound to affect its immunogenecity could be by mutation of a protein allergen in its immunoglobulin-specific epitopes.
- the location of these epitopes can be determined by several techniques such as those disclosed by WO 92/10755 (by U. L ⁇ vborg), by Walshet et al, J. Immunol. Methods, vol. 121, 1275-280, (1989), and by Schoofs et al. J. Immunol, vol. 140, 611-616, (1987).
- a prefened method for identification of epitopes is by screening a random peptide library with antibodies (e.g.
- IgE or IgG antibodies selecting high-binding peptides, obtaining the sequence these, and aligning the high-binding peptide sequences to identify a consensus sequence.
- These consensus sequences are compared with the sequence and 3D structure of a relevant protein, which is desired to mutate for reduction of immunogenicity, in order to identify the linear and structural epitopes of that protein.
- the identification of epitope(s) may be achieved by screening of phage display libraries.
- the principle behind phage display is that a heterologous DNA sequence can be inserted in the gene coding for a coat protein of the phage. The phage will make and display the hybrid protein on its surface where it can interact with specific target agents.
- Such target agent may be antigen-specific antibodies. It is therefore possible to select specific phages that display antibody-binding peptide sequences.
- the displayed peptides can be of predetermined lengths, for example 9 amino acids long, with randomized sequences, resulting in a random peptide display package library. Thus, by screening for antibody binding, one can isolate the peptide sequences that have the highest affinity for the particular antibody used.
- the peptides of the hybrid proteins of the specific phages which bind protein- specific antibodies define the epitopes of that particular protein.
- the conesponding residues of the parent protein can be found by aligning the selected peptide sequences resulting in epitope patterns, and compare these with the amino acid sequence and 3-dimensional structure of the parent protein.
- a protein variant exhibiting a reduced immunogenicity may be produced by changing the identified epitope pattern of the parent protein by genetic engineering of a DNA sequence encoding the parent protein. It is commonly found, that amino acids sunounding B- and T-cell epitopes can affect binding of the antibodies or T-cell receptors to the antigen. To anticipate this possibility, an epitope area was defined on the 3-dimensional structure of the protein of interest, and genetic engineering of any amino acid within the epitope area is considered to be within the scope of using the identified epitope to generate variants with low immunogenecity.
- amino acid residues may be substituted, added or deleted, these amino acids preferably being located in different epitope areas.
- a diversified library can be established by a range of techniques known to the person skilled in the art (Reetz MT; Jaeger KE, in 'Biocatalysis - from Discovery to Application' edited by Fessner WD, Vol. 200, pp. 31-57 (1999); Stemmer, Nature, vol. 370, p.389-391, 1994; Zhao and Arnold, Proc. Natl. Acad. Sci., USA, vol. 94, pp. 7997-8000, 1997; or Yano et al., Proc. Natl. Acad. Sci., USA, vol. 95, pp 5511-5515, 1998).
- oligonucleotide primers which are synthesized using a mixture of nucleotides for certain positions.
- the mixtures of oligonucleotides used within each triplet can be designed such that the s conesponding amino acid of the mutated gene product is randomized within some predetermined distribution function. Algorithms exist, which facilitate this design (Jensen LJ et al., Nucleic Acids Research, Vol. 26(3), 697-702 (1998)).
- substitutions are found by a method comprising the following steps: 1) a o range of substitutions, additions, and/or deletions are listed encompassing several epitope areas, 2) a library is designed which introduces a randomized subset of these changes in the amino acid sequence into the target gene, e.g. by spiked mutagenesis, 3) the library is expressed, and prefened variants are selected.
- this method is supplemented with additional rounds of screening and/or family shuffling of hits from the first 5 round of screening (J.E. Ness, et al, Nature Biotechnology, vol. 17, pp. 893-896, 1999) and/or combination with other methods of reducing immunogenicity by genetic means (such as that disclosed in WO92/10755).
- the library may be designed, such that at least one amino acid of the epitope area is o substituted. In a prefened embodiment at least one amino acid of the epitope itself is changed.
- the library may be biased such that towards introducing an amino acid of different size, hydrophilicity, and/or polarity relative to the original one of the 'protein backbone'. For example changing a small amino acid to a large amino acid, a hydrophilic amino acid to a hydrophobic amino acid, a polar amino acid to a non-polar amino acid or a basic to an acidic amino acid. Other changes may be the addition or deletion of at least one amino acid of the epitope area, preferably deleting an anchor amino acid. Furthermore, substituting some amino acids and deleting or adding others may change an epitope.
- the library is designed, such that recognition sites for post- translational modifications are introduced in the epitope areas, and the library is expressed in a suitable host organism capable of the conesponding post-translational modification.
- These post-translational modifications may serve to shield the epitope and hence lower the immunogenicity of the protein variant relative to the protein backbone.
- Post-translational modifications include glycosylation, phosphorylation, N-terminal processing, acylation, ribosylation and sulfatation. A good example is N-glycosylation.
- N-glycosylation is found at sites of the sequence Asn-Xaa-Ser, Asn-Xaa-Thr, or Asn-Xaa-Cys, in which neither the Xaa residue nor the amino acid following the tri -peptide consensus sequence is a pro line (T. E. Creighton, 'Proteins - Structures and Molecular Properties, 2nd edition, W.H. Freeman and Co., New York, 1993, pp. 91-93). It is thus desirable to introduce such recognition sites in the sequence of the backbone protein.
- the specific nature of the glycosyl chain of the glycosylated protein variant may be linear or branched depending on the protein and the host cells.
- the protein variant can then be conjugated to activated polymers Which amino acids to substitute and/or insert depends in principle on the coupling chemistry to be applied.
- the chemistry for preparation of covalent bioconjugates can be found in "Bioconjugate Techniques", Hermanson, G.T. (1996), Academic Press Inc., which is hereby incorporated as reference (see below).
- Diversity in the protein variant library can be generated at the DNA triplet level, such that individual codons are variegated e.g. by using primers of partially randomized sequence for a PCR reaction. Further, several techniques have been described, by which one can create a library with such diversity at several locations in the gene, which are too far apart to be covered by a single (spiked) oligonucleotide primer. These techniques include the use of in vivo recombination of the individually diversified gene segments as described in WO 97/07205 on page 3, line 8 to 29 or by using DNA shuffling techniques to create a library of full length genes that combine several gene segments each of which are diversified e.g. by spiked mutagenesis (Stemmer, Nature 370, pp.
- Host cells culturing, and sampling
- any number of host cells can be used to perform the method of the invention.
- the host cells are of microbial origin, preferably bacterial, yeast, or fungal.
- the host cells are chosen from the group consisting of Escherichia coli, Bacillus subtilis, Bacillus licheniformis, Bacillus clausii, Aspergillus niger, Aspergillus oryzae, Aspergillus nidulans, and Saccharomyces cerevisiae.
- the diversified library is prepared as a DNA library of genes encoding variants of the relevant protein backbone.
- This DNA library is then transformed into the host cells by using any of the techniques known in the art and described above.
- the number of cells per well is optimized such that the highest density of anay positions occupied by exactly one cell is obtained.
- the number of cells at each position can be controlled by dilution. The dilution that most closely approximates one cell per position is often termed 'the limiting dilution'.
- Steps (iii) through (vi) of the present invention can be performed in many ways.
- the culturing and the sampling of host cells takes place using a spatial anay.
- the spatial anay can take on any physical form whatsoever, that enables the culturing or assaying of several samples at once, without one sample contaminating another.
- Examples of prefened spatial anays are different kinds of microtiter-plates with any number of wells, such as 96 or 384, and of any kind of material, as well as positions in a High Performance Liquid Chromatography (HPLC) autosampler device. Any kind of physical anangement which allows the unambiguous identification of the samples by a number or a position in the anay.
- HPLC High Performance Liquid Chromatography
- a prefened embodiment relates to a method, wherein the spatial anay of is a microtiter plate, a solid surface or a textile surface.
- a way of carrying out step (iv) of the first aspect of the invention could be to take a sample 5 from each position of the spatial anay, e.g. from a supernatant or cell culture, and transfer this to another spatial anay for further testing or assaying for production of the molecule of interest.
- the second spatial anay may be identical to the first one used in the specific method, but may also be of any other kind that fulfills the above mentioned criteria, such as a microtiter plate, a solid surface, a textile, any material etc.
- a prefened embodiment relates to a method, wherein after step (iv) a sample is transfened from each position in the spatial anay to a position in a second spatial anay which is then used onwards in the method, preferably the second anay is a microtiter plate.
- the individual cells/clones are grown within microenvironments in each position of the spatial anay.
- microenvironments are initally s sterile beads or balls of any material that allows growth of the clone, preferably beads comprising agarose, alginate, polysaccharide, carbohydrate, alginate, canageenan, chitosan, cellulose, pectin, dextran or polyacrylamide, all allowing diffusion to/from each micro- environment.
- a prefened embodiment relates to a method of the first aspect, wherein each position in the spatial anay is occupied by a bead comprising one cell, preferably the bead is an agarose-bead.
- the assaying of the invention for production of the molecule of interest can be done in a great 5 number of ways.
- some kinds of spatial anays like microtiter plates, can sometimes be assayed directly, or samples can be taken from each position and tranfened to another spatial anay for assaying according to step (v) and (vi) or according to any of a number of techniques well known in the art, such as enzyme activity, receptor binding, and many others; common to these assays is that a measurement is taken of a detectable property o e.g. fluorescence, luminescence, absorption.
- the method of the invention can be performed with any number of these assays, and consequently the molecules assayed for in step (v) and (vi) can be obtained in a number of ways, depending on the host cell construct.
- the molecule of interest may be secreted by the host cell into the supernatant, or the molecule may remain intracellular, in which case lysis of the host cells may release the molecule.
- a prefened embodiment relates to a method, wherein the molecule of interest in (v) and (vi) is assayed in either whole broth, supernatant of cells that secrete the molecule, a lysate of cells that produce the molecule, or is assayed while still inside cells that produce the molecule.
- the sample is purified by a size separation process, such as the membrane processes filtration and dialysis, prior to the analyses of step (v) and (vi).
- a size separation process such as the membrane processes filtration and dialysis, prior to the analyses of step (v) and (vi).
- This size separation can be achieved, e.g. by using microtiter well inserts (e.g. Nunc TC), vacuum filtration microtiter plates (e.g. Qiagen, QIAwell) or pumping or centrifugation devices.
- the host cells may be first selected based on a functional screen, and then re-cultured, sampled, and analysed for antibody binding capacity and functionality.
- the initial functional selection can be done using an agar plate assay to analyse colonies from the transformed cells and select for those resulting in halo formation (see for example Ness, J.E:, et al., Nature Biotechn., 17, pp 893-896, 1999).
- host cells expressing functional protein variants may be selected, preferably by automated colony picking, before the antibody binding capacity of the protein variant is assayed. This mode can increase the throughput and quality of the screening setup in cases where the antibody binding capacity assay is slower, more cumbersome, more expensive, or less accurate than the functionality assay.
- protein variant is exposed to adverse conditions prior to analyzing the functionality, in order to gauge also the stability of the protein variant.
- the functionality is determined twice with the protein variant being exposed to adverse conditions for a period of time in between the two determinations of functionality. From these measurements the stability of the protein variant at the particular set of adverse conditions can be determined.
- These adverse conditions could be characterized by increased temperature, increased or decreased pH, the presence of certain metal ions or of metal chelators such as EDTA. They could also be the presence of surfactant molecules, such as those used in detergents, skin cream formulations, hand dish washing compositions or other applications; the presence of proteases or other degradative enzymes; or they could be the presence of enzyme inhibitors.
- the protein variant of the sample (step iv) is immobilized to a solid material.
- a solid material could be the well surface of a microtiterplate, suspended beads, the pin of dipstick devices or others.
- Immobilization can be either nonspecific by hydrophobic interactions, ionic interactions, or chemical coupling, or it can be more specific, such as non-covalent coupling to immobilized binding partners such as antibodies, enzyme inhibitors, or substrate mimics.
- the protein backbone has been modified genetically with an N-terminal or C-terminal extension comprising a high affinity tag for a compound, which can be immobilized.
- a high affinity tag for a compound which can be immobilized.
- An example is a polyhistidine tag, which binds with high affinity to Nickel, which in turn has been bound to an acid such as nitrilotriacetic acid which has been immobilized to the solid phase.
- affinity tags are the cellulose binding domains of bacteria and fungi, which bind with high affinity to cellulose (such as Avicell), also when fused onto heterogeneous proteins, calmodulin binding domains, S-tag or FLAG peptides that bind to specific antibodies etc.
- the binding to the solid phase is reversible, such that the protein variant can be eluted into solution when exposed to certain conditions such as e.g. high salt concentrations, high pH, or high competitor concentration.
- efficient HTS is achieved by devicing a simple yet accurate determination of the amount of a specific active molecule produced by the individual clone.
- concentration of the specific molecule in a position of the spatial anay has been determined, this information can be used to determine the specific activity from the total activity determined in that position; alternatively, the information can be used to adjust the input of the molecule into the activity assay.
- a second solution is to use immobilization of the protein variant as a means to dose the subsequent assays with a known and/or constant amount of protein variant. This requires that the sample contain more protein variant than the binding capacity of the immobilization step.
- the assay can be configured in such a way that it is insensitive to the concentration of the molecule.
- the protein variants are being modified by chemical conjugation prior to the analyses of step (v) and (vi).
- the protein variant needs to be incubate with an active or activated polymer and subsequently separated from the unreacted polymer. This can conveniently be done using the immobilized protein variants, which can easily be exposed to different reaction environments and washes.
- polymeric molecules are to be conjugated with the polypeptide in question and the polymeric molecules are not active they must be activated by the use of a suitable technique. It is also contemplated according to the invention to couple the polymeric molecules to the polypeptide through a linker. Suitable linkers are well-known to the skilled person.
- Some of the methods concern activation of insoluble polymers but are also applicable to activation of soluble polymers e.g. periodate, trichlorotriazine, sulfonylhalides, divinylsulfone, carbodiimide etc.
- the functional groups being amino, hydroxyl, thiol, carboxyl, aldehyde or sulfydryl on the polymer and the chosen attachment group on the protein must be considered in choosing the activation and conjugation chemistry which normally consist of i) activation of polymer, ii) conjugation, and iii) blocking of residual active groups.
- Coupling polymeric molecules to the free acid groups of polypeptides may be performed with the aid of diimide and for example amino-PEG or hydrazino-PEG (Pollak et al., (1976), J.
- Coupling polymeric molecules to hydroxy groups is generally very difficult as it must be performed in water. Usually hydrolysis predominates over reaction with hydroxyl groups.
- Coupling polymeric molecules to free sulfhydryl groups can be achieved with special groups like maleimido or the ortho-pyridyl disulfide. Also vinylsulfone (US patent no. 5,414,135,
- Accessible Arginine residues in the polypeptide chain may be targeted by groups comprising two vicinal carbonyl groups.
- PEG into good leaving groups (sulfonates) that, when reacted with nucleophiles like amino groups in polypeptides allow stable linkages to be formed between polymer and polypeptide.
- the reaction conditions are in 5 general mild (neutral or slightly alkaline pH, to avoid denaturation and little or no disruption of activity), and satisfy the non-destructive requirements to the polypeptide.
- Tosylate is more reactive than the mesylate but also less stable decomposing into PEG, dioxane, and sulfonic acid (Zalipsky, (1995), Bioconjugate Chem., 6, 150-165).
- Epoxides may also been used for creating amine bonds but are much less reactive than the abovementioned lo groups.
- the derivatives are usually made by reacting the chloroformate with the desired leaving group. All these groups give rise to carbamate linkages to the peptide. Furthermore, isocyanates and isothiocyanates may be employed, yielding ureas and thioureas,
- Amides may be obtained from PEG acids using the same leaving groups as mentioned above and cyclic imid thrones (US patent no. 5,349,001, (1994), Greenwald et al.). The reactivity of these compounds are very high but may make the hydrolysis to fast. PEG succinate made from reaction with succinic anhydride can also be used. The hereby
- ester group make the conjugate much more susceptible to hydrolysis (US patent no. 5,122,614, (1992), Zalipsky). This group may be activated with N-hydroxy succinimide. Furthermore, a special linker can be introduced. The most well studied being cyanuric chloride (Abuchowski et al, (1977), J. Biol. Chem., 252, 3578-3581; US patent no. 4,179,337, (1979), Davis et al.; Shafer et al., (1986), J. Polym. Sci. Polym. Chem. Ed., 24,
- Coupling of PEG to an aromatic amine followed by diazotation yields a very reactive diazonium salt, which can be reacted with a peptide in situ.
- An amide linkage may also be obtained by reacting an azlactone derivative of PEG (US patent no. 5,321,095, (1994), Greenwald, R. B.) thus introducing an additional amide linkage.
- peptides do not comprise many Lysines it may be advantageous to attach more than one PEG to the same Lysine. This can be done e.g. by the use of l,3-diamino-2-propanol. PEGs may also be attached to the amino-groups of the enzyme with carbamate linkages (WO 95/11924, Greenwald et al.). Lysine residues may also be used as the backbone.
- the coupling technique used in the examples is the N-succinimidyl carbonate conjugation technique descried in WO 90/13590 (Enzon).
- the activated polymer is methyl-PEG which has been activated by N-succinimidyl carbonate as described WO 90/13590.
- the coupling can be carried out at alkaline conditions in high yields.
- a methyl- PEG 350 could be activated with N-succinimidyl carbonate and incubated with protein variant at a molar ratio of more than 5 calculated as equivalents of activated PEG divided by moles of lysines in the protein of interest.
- the PEG:protein ratio should be optimized such that the PEG concentration is low enough for the buffer capacity to maintain alkaline pH throughout the reaction; while the PEG concentration is still high enough to ensure sufficient degree of modification of the protein. Further, it is important that the activated PEG is kept at conditions that prevent hydrolysis (i.e. dissolved in acid or solvents) and diluted directly into the alkaline reaction buffer. It is essential that primary amines are not present other than those occurring in the lysine residues of the protein. This can be secured by washing thoroughly in borate buffer. The reaction is stopped by separating the fluid phase containing unreacted PEG from the solid phase containing protein and derivatized protein. Optionally, the solid phase can then be washed with tris buffer, to block any unreacted sites on PEG chains that might still be present. Determining antibody binding capacity:
- a sample is analysed to determine the antibody binding capacity of its variant protein.
- the antibodies used for this analysis can be in the form of serum isolated from an animal such as rabbit, mouse, rat, guinea pig, sheep etc., which has previously been exposed to the protein backbone.
- the serum can be human serum from a volunteer who is known to have a history of exposure to the protein backbone.
- serum samples are pooled from several animal or human donors to achieve an individual-independent result of the screening.
- the serum can be purified (e.g.
- the IgE-purification method is prefened, and even more preferable is an embodiment in which the experimental animal has been intratracheally exposed to the protein backbone and has been shown to develop antigen-specific IgE antibodies, or if the human volunteers have a history of allergenic sensitisation to the protein backbone.
- IgG, IgE, or other immunoglobulin fraction it is preferable to increase the specificity of the antibody binding analysis by further purification.
- This can be in the form of affinity purification using a solid phase of immobilized antigen (e.g. the protein backbone).
- Immobilization can be by chemical conjugation, specific binding to a ligand or an antibody, or by binding through a fusion tag such as a polyhistidine tage, a FLAG peptide or the like.
- Antibodies are bound and eluted (e.g. using 1 M propionic acid) resulting in a "specific polyclonal antibody preparation" (see Arvieuz and Williams: Immunoaffinity
- protease antigen it may be desirable to reduce or eliminate protease activity during the antibody purification step. This can be done by several methods, including binding to an immobilized inhibitor (such as bacitracin in the case of subtilisins), using a chemically inactivated protease backbone (e.g. by PMSF treatment), or by using a mutated protease (e.g. 5 by converting the catalytically active serine to an alanine). In the latter case, the antibodies are raised against the mutated version of the protein backbone, while the library is diversified using the active protein backbone as a template, and this embodiment of the method is also considered an aspect of the present invention.
- an immobilized inhibitor such as bacitracin in the case of subtilisins
- PMSF treatment a chemically inactivated protease backbone
- mutated protease e.g. 5 by converting the catalytically active serine to an alanine
- the antibodies can be purified using an epitope-specific ligand.
- this can be in the form of a peptide fragment of the protein backbone, while in the case of a structural epitope this could be in the form of a peptide (or peptide phage membrane protein fusion) which has been isolated by a peptide display library using antigen-specific polyclonal antibodies, as described above.
- Such purification schemes lead to s "monospecific antibodies" which are useful for the cunent invention.
- the antibodies can be monoclonal antibodies, each of which are considered a subgroup of "monospecific antibodies" as they have only one binding speicificity.
- the hybridoma clones can be selected based on specificity for the entire protein o backbone or for specificity for an epitope (as described above). In either case, the epitope specificity can be assessed for instance by standard immunoassays using the isolated antibody- binding peptides in order to assign the specificity of each hybridoma clone to a particular epitope pattern. This way, one can create libraries with variation in a single epitope and assay it with one or more monoclonal antibodies specific for that particular epitope in order to get a 5 very specific response.
- the antibodies can be labelled to allow detection.
- secondary antibodies directed against the primary antibodies can be labelled.
- the label can be a chemically bound compound such as peroxidase, o streptavidin, alkaline phosphatase, fluorescent or luminescent compounds or others, or it can be a fusion tag such as a polyhistidine tag, a FLAG peptide, an S-tag or other fusion peptides for which there are or can be made specific antibodies.
- This background interference can be reduced by thorough purification of the protein backbone prior to sensitisation of the test animal, or by several other methods.
- One such method is to purify the antibodies on a column with immobilized 'cellular impurities' obtained by culturing a strain of host cells which have not been transformed or which have been transformed with a vector that does not contain any protein backbone or protein variant, as described (Naver and L ⁇ vborg, Scand. J. Immunol., 41, pp. 443-448, 1995).
- Another method to reduce background interference is to raise the antibodies against a protein backbone, which has been expressed in an organism or strain different from the host cells used for expression of the diversified libraries.
- a third method is to use immobilization of the protein variants (as described above) to facilitate removal of cell culture supernatant impurities.
- the antibody binding assay is multivalent in nature, i.e. it depends on multivalent or bivalent interactions between antibody and antigen.
- assay formats are agglutination assays and assays based on 'passive immunization' of effector cells, e.g. mast cells or basophiles, with IgE antibodies and detection of cell-specific responses to antigen-induced aggregation of the cell-surface bound IgE molecules (Skov, PS et al, Pediatr. Allergy Immunol., 8, pp. -156-158, 1995; Diamant and Pratkar, Int. Archs. Allergy appl. Immun. 67, pp. 13-17, 1982).
- This aspect can be advantageous when several epitopes are diversified in the same library.
- the antibody binding capacity is determined in step (v).
- an antibody binding analysis that requires few dosages, preferably only one dosage of protein variant, and which in other ways is designed to give the highest likelihood of identifying low immunogenic protein variants.
- the binding can be determined using a labelled primary antigen-specific antibody or alternatively by using a primary antigen-specific antibody and a labelled secondary species- specific anti-Ig antibody. Immobilization of the protein variant makes it easier to change the reaction medium several times, introduce washing steps etc.
- the antibody-binding can be determined using competitive antigen (e.g. protein backbone) and o labelled antibodies in the solution.
- the protein variants may be analysed for antibody binding capacity directly, or they may have been immobilized to a solid phase and then eluted (as described above). In either case, the s antibody binding capacity is analysed with the protein variant in solution.
- the antibodies have been coated onto a solid phase (such as the surface of wells in a microtiter plate.
- a solid phase such as the surface of wells in a microtiter plate.
- protein variants bind to (or be captured by) the immobilized antibodies at the surface, and the supernatant contain only antigens that do not bind well to the o antibodies (provided that the relative amounts of coated antibody and added protein variant have been adjusted such that the protein backbone when added in similar amounts as the protein variant, binds fully to the coated antibodies).
- the supernatant is withdrawn and assayed for functionality to determine whether functional protein variants have reduced antibody-binding capacity. In this mode, the analysis of 5 antibody binding and protein variant functionality are combined to give a single read-out.
- the conditions can be adjusted to lower the affinity between antibody and protein backbone in order to allow protein variants with moderately reduced antibody binding capacity to be free in solution.
- the affinity can be lowered for instance by adjusting the salt concentration and or pH or by carrying out the incubation in the presence of surface active 0 ingredients or in the presence of competitive modified protein backbone, which has been inactivated (chemically or genetically, as described above for proteases) in order to give no signal in the functionality assay.
- the antibody-coated surfaces are incubated with protein variant and labelled competitive antigen in a ratio that ensures that no or minimal amounts of labelled competitor are bound to the surface when incubated with the protein backbone. After incubation, the supernatant is removed and the amount of bound competitor is determined.
- the advantage of this approach is that a protein variant that has had an epitope eliminated will be less likely to give a false positive signal by binding through a different epitope. If the protein variant has had an epitope eliminated the subset of antibodies that are specific for this epitope will be unoccupied and allow binding of labelled competitor to the surface, even when the labelled competitor is present in far lower concentration than the protein variant.
- protease assays WO99/34011, Genencor International; J.E. Ness, et al, Nature Biotechn., 17, pp. 893-896, 1999
- oxidoreductase assays Choerry et al., Nature Biotechn., 17 , pp. 379-384, 1999
- assays for several other enzymes WO99/45143, Novo Nordisk.
- protein variants have been selected based on the methods described in this invention, it is desirable to confirm their antibody binding capacity, functionality, and immunogenicity using a purified preparation.
- a selected clone should be re-isolated by conventional microbiological techniques and its characteristics tested in the same assay system, then (if results are confirmed) the protein variant of interest should be expressed in larger scale, purified by conventional techniques, and the reduced antibody binding capacity and the functionality should be examined in detail using dose-response curves and e.g. competitive ELISA (C-ELISA).
- C-ELISA competitive ELISA
- the potentially reduced allergenicity (which is likely, but not necessarily true for a variant w. low antibody binding) should be tested in in vivo or in vitro model systems: e.g. an in vitro assays for immunogenicity such as assays based on cytokine expression profiles or other proliferation or differentiation responses of epithelial and other cells incl. B-cells and T-cells.
- animal models for testing allergenicity should be set up to test a limited number of protein variants that show desired characteristics in vitro.
- Useful animal models include the guinea pig intratracheal model (GPIT) (Ritz, et al. Fund. Appl. Toxicol., 21, pp.
- mouse subcutaneous WO 98/30682, Novo Nordisk
- rat-IT rat intratracheal
- MINT mouse intranasal
- the immunogenicity of the protein variant is measured in animal tests, wherein the animals are immunised with the protein variant and the immune response is measured. Specifically, it is of interest to determine the allergenicity of the protein variants by repeatedly exposing the animals to the protein variant by the intratracheal route and following the specific IgG and IgE titers. Alternatively, the mouse intranasal (MINT) test can be used to assess the allergenicity of protein variants.
- the allergenicity is reduced at least 10 times as compared to the allergenicity of the parent protein, preferably 50 times reduced, more preferably 100 times.
- the present inventors have demonstrated that the performance in a competitive ELISA conelates closely to the immunogenic responses measured in animal tests.
- the IgE binding capacity of the protein variant must be reduced to at least below 75 %, preferably below 50 % of the IgE binding capacity of the parent protein as measured by the performance in competitive IgE ELISA, given the value for the IgE binding capacity of the parent protein is set to 100 %.
- Horse Radish Peroxidase labelled anti-rabbit-Ig (Dako, DK, P217; dilution 1 :1000).
- Rabbit anti-Savinase polyclonal IgG prepared by conventional means.
- Cyanuric chloride (Aldrich), Acetone (Merck),Tween 20 (Merck), Skim Milk powder (Difco), H2SO4 (Merck), OPD: o-phenylene-diamine: (Kementec cat no. 4260), H2O2, 30% (Merck).
- Carbonate buffer (0.1 M, pH 10): Na2CO3 10.60 g/L.
- PBS pH 7.2: NaCl 8.00 g/L; KC1 0.20 g/L; K2HPO4 1.04 g/L; KH2PO4 0.32 g/L.
- Washing buffer PBS, 0.05% (v/v) Tween 20.
- Blocking buffer PBS, 2% (wt/v) Skim Milk powder.
- Dilution buffer PBS, 0.05% (v/v) Tween 20, 0.5% (wt/v) Skim Milk powder.
- a fresh stock solution of 10 mg/ml cyanuric chloride in acetone is diluted into PBS, while stirring, to a final concentration of 1 mg/ml and immediately aliquoted into CovaLink NH2 plates (100 microliter per well) and incubated for 5 minutes at room temperature. After three washes with PBS, the plates are dryed at 50°C for 30 minutes, sealed with sealing tape, and stored in plastic bags at room temperature for up to 3 weeks.
- Immobilization of antibody/competitive antigen Activated CovaLink NH2 plates are coated overnight at 4 °C with 100 microliter of the desired protein (5 micro gram/ml) in PBS followed by 30 min incubation with blocking buffer at room temperature and four washes in PBS-tween.
- Proteases cleave the bond between the peptide and p-nitroaniline to give a visible yellow colour absorbing at 405 nm.
- 100 mg suc-AAPF-pNa is dissolved into 1 ml dimethyl sulfoxide (DMSO). 100 microliter of this is diluted into 10 ml with Britton and Robinson o buffer, pH 8.3, and used as substrate for the protease. Reaction is detected kinetically in a spectrophotometer.
- DMSO dimethyl sulfoxide
- the supernatant from the culture medium is diluted 200-fold in the reaction mixture, which s contains 5 microg/ml BODIPY FL-casein (Molecular Probes), 1 mM CaCl , and 50 mM Tris- HCL, pH 7,5. After 60 min. incubation at room temperature, the plates are read at 520 nm with excitation at 485 nm using a FLUOstar (BMG Technologies).
- Example 1 Capturing antigen.
- a protease antigen is captured on immobilized antibodies and the non-captured 5 fraction is assayed for functionality. This is performed on CovaLink NH2 plates coated with rat anti-Savinase IgE.
- Protease libraries producing epitopee variants were established. These variants might or might not be His-tagged. Screening is directly on bacterial culture media, using covalink plates 0 coated with mouse anti-rat IgE monoclonal antibodies satured with anti-savinase specific rat IgE. The amounts of bound wild type antigen was determined with a anti-wild type polyclonal rabbit antiserum.
- a 'backbone protease' protease inhibitor is immobilized in the wells and incubated with an excess of the protein variant and labelled antibodies. The level of bound antibodies is determined.
- 25 microliter sample and 25 microliter anti-Savinase antibody are added to the coated well and incubated at room temperature (30 min). The supernatant is removed and the wells are washed three times in PBS-tween.
- a separate sample is analysed for functionality and the two values are compared.
- Desired protein variants show a high level of bound antibody and at the same time a level of functionality similar to the 'backbone protein'.
- Example 3 Immobilization of His-tagged proteases.
- the DNA sequence encoding the protease Savinase® (Novo Nordisk A/S, Denmark) is translationally fused to a sequence encoding a polyhistidine tag (His6) and libraries of Savinase®-His6 variants are produced and introduced into Bacillus. After standard culturing, a limited number of Savinase® enzymes of each variant (about 10% of what is secreted by Bacillus carrying the wildtype Savinase gene) are immobilized in the wells of Ni-NTA microtiter plates.
- the unbound fraction including cells and excess Savinase® is removed, and the plate washed once or twice in a buffer containing 5-20 mM Imidazole.
- the immobilized Savinase can now be assayed for antibody binding capacity directly, or used modified with activated PEG and washed prior to analysis.
- the His-tagged Savinase® variants are released from the solid support by the addition of 250 mM Imidazole, and aliquots of the supematants from each well are sampled for antibody binding and functionality assays as described in the previous examples.
Landscapes
- Life Sciences & Earth Sciences (AREA)
- Health & Medical Sciences (AREA)
- Engineering & Computer Science (AREA)
- Molecular Biology (AREA)
- Chemical & Material Sciences (AREA)
- Immunology (AREA)
- Physics & Mathematics (AREA)
- Urology & Nephrology (AREA)
- Biomedical Technology (AREA)
- Hematology (AREA)
- Biochemistry (AREA)
- Medicinal Chemistry (AREA)
- General Health & Medical Sciences (AREA)
- Bioinformatics & Cheminformatics (AREA)
- Bioinformatics & Computational Biology (AREA)
- Biophysics (AREA)
- Microbiology (AREA)
- Pathology (AREA)
- Biotechnology (AREA)
- Cell Biology (AREA)
- Proteomics, Peptides & Aminoacids (AREA)
- General Physics & Mathematics (AREA)
- Analytical Chemistry (AREA)
- Food Science & Technology (AREA)
- Organic Chemistry (AREA)
- Chemical Kinetics & Catalysis (AREA)
- General Chemical & Material Sciences (AREA)
- Micro-Organisms Or Cultivation Processes Thereof (AREA)
- Preparation Of Compounds By Using Micro-Organisms (AREA)
- Peptides Or Proteins (AREA)
Abstract
The present invention is to provide a method to perform assays that efficiently and accurately can screen large numbers of cell populations producing variants of a molecule of interest. The present invention relates more specifically to a method for screening a library of protein variants for functional variants with reduced antibody binding capacity, comprising the steps of: (i) generating a diversified library of protein variants starting from a relevant protein backbone, (ii) transforming the library into suitable host cells, (iii) culturing host cells, (iv) sampling each cell culture, (v) analysing a sample by determining the antibody binding capacity of the variant protein, (vi) analysing a sample by determining the functionality of the variant protein.
Description
High Throughput Screening (HTS) assays
Field of invention
This invention relates to methods for screening libraries of protein variants for those having reduced immunogenicity as compared to a protein backbone. More specifically, the present invention provides a method for identifying protein variants with reduced immunogenicity in an efficient manner.
Background of the invention
An increasing number of proteins, including enzymes, are being produced industrially, for use in various industries, housekeeping and medicine. Being proteins they are likely to stimulate an immunological response in man and animals, including an allergic response.
In the present context the terms allergic response, allergy, allergenic and allergenicity are used according to their usual definitions, i.e. to describe the reaction due to immune responses wherein the antibody most often is IgE, less often IgG4 and diseases due to this immune response. Allergic diseases include urticaria, hay-fever, asthma, and atopic dermatitis. They may even evolve into an anaphylactic shock.
Prevention of allergy in susceptible individuals is therefore a research area of great importance. Depending on the application, individuals get sensitized to the respective allergens by inhalation, direct contact with skin and eyes, or ingestion. The general mechanism behind an allergic response is divided in a sensitization phase and a symptomatic phase. The sensitization phase involves a first exposure of an individual to an allergen. This event activates specific T- and B-lymphocytes, and leads to the production of allergen-specific IgE antibodies (in the present context the antibodies are denoted as usual, i.e. immunoglobulin E is IgE etc.). These IgE antibodies eventually facilitate allergen capturing and presentation to
T-lymphocytes at the onset of the symptomatic phase. This phase is initiated by a second exposure to the same or a resembling antigen. The specific IgE antibodies bind to the specific IgE receptors on mast cells and basophils, among others, in a mode that allows antigen binding to the cell-bound IgE antibodies. The polyclonal nature of this process results in bridging and clustering of the IgE receptors, and subsequently in the activation of mast cells and basophiles. This activation triggers the release of various chemical mediators involved in the early as well as late phase reactions of the symptomatic phase of allergy.
Various attempts to reduce the immunogenicity of polypeptides and proteins have been conducted. It has been found that small changes in an epitope may affect the binding to an antibody. This may result in a reduced importance of such an epitope, maybe converting it from a high affinity to a low affinity epitope, or maybe even resulting in epitope loss, i.e. that the epitope cannot sufficiently bind an antibody to elicit an immunogenic response.
Technologies such as DNA shuffling, random DNA mutagenesis and in vivo recombination have allowed the generation of enormous populations of variant cells that produce variants of a certain protein. In addition, it has become possible in recombinant host strains to establish large libraries of natural enzymes cloned from other organisms. Together these technologies have created a need for assays that efficiently and accurately can screen large numbers of variants, high throughput screening (HTS) assays.
Most HTS methods are designed to detect an improved functionality of the expressed protein variants. In this case, however, we are interested in isolating variants with reduced antibody binding capacity, and hence we risk to select variants with debilitated functionality (which are also likely to bind antibodies poorly) or variants which expresses and/or secretes poorly and hence show very low antibody binding simply because they are not present in the supernatant. To overcome this limitation it is necessary to deviate from the design of most present HTS methods and introduce a dual assay system in which both antibody binding capacity and functionality are determined without compromising the high throughput capability necessary to benefit the diversity of diversified libraries.
The prior art, different publications have disclosed aspects that are useful for creating protein variants with altered antibody binding capacity, but none have disclosed a method for screening a diversified library expressed in host cells to search for functional protein variants with reduced antibody binding capacity.
For instance, WO 99/447680 discloses the modification of B-cell epitopes by protein engineering. However, the method is based on crystal structures of Fab-antigen complexes, and B-cell epitopes are defined as "a section of the surface of the antigen comprising 15-25 amino acid residues, which are within a distance from the atoms of the antibody enabling direct interaction" (p.3). This publication does not show how one selects which Fab fragment to use (e.g. to target the most dominant allergy epitopes) or how one selects the substitutions to be made. Further, their method cannot be used in the absence of such crystallographic data, which is very cumbersome, sometimes impossible, to obtain - especially since one would need a separate crystal structure for each epitope to be changed.
Slootstra et al; Molecular Diversity, 2, pp. 156-164, 1996 disclose the screening of a semirandom library of peptides for their binding properties to three monoclonal antibodies by immobilizing the peptides on polyethylene pins and binding a dilution series of each antibody to the pins. In this reference, all peptides are prepared by chemical synthesis; hence, it does not disclose any method to overcome background problems from using gene-encoded polypeptides expressed in a microbial host. Further, the antibody binding assays are based on 10-step dilution series for each antibody, meaning that 30 separate assays are necessary to evaluate each test compound. This makes the disclosed methods insufficient for high- throughput screening.
WO 97/30150 discloses the construction and expression of diversified libraries of a myelin basic protein (MBP) and the analysis of these variants by testing their T-cell antagonizing activity. This reference, however, only tests for 'trans-dominant effects' (p. 17) in which a single peptide harboring a productive mutation will show up even in the presence of 10 unchanged peptide fragments. This means that dysfunctional or poorly expressed protein variants will show no response and hence, that the teachings of this reference cannot be used
for devising an assay to identify protein variants with reduced antibody binding capacity and retained functionality.
Below we describe a method of performing an assay that has been specifically developed for the screening of large populations of clones producing variants of a given protein.
The methods described below allow screening library of protein variants for functional variants with reduced antibody binding capacity.
Summary of the invention
The problem to be solved by the present invention is to provide a method to perform assays that efficiently and accurately can screen large numbers of cell populations producing variants of a molecule of interest.
In a first aspect the invention relates to a method for high throughput screening (HTS) of a large population of host cells for production of a molecule of interest.
Specifically, the invention relates to a method for screening a library of protein variants for functional variants with reduced antibody binding capacity, comprising the steps of:
(i) generating a diversified library of protein variants starting from a relevant protein backbone,
(ii) transforming the library into suitable host cells,
(iii) culturing host cells,
(iv) sampling each cell culture,
(v) analysing a sample by determining the antibody binding capacity of the variant protein,
(vi) analysing a sample by determining the functionality of the variant protein.
Definitions
Prior to a discussion of the detailed embodiments of the invention, a definition of specific terms related to the main aspects of the invention is provided. 0
In accordance with the present invention there may be employed conventional molecular biology, microbiology, and recombinant DNA techniques within the skill of the art. Such techniques are explained fully in the literature. See, e.g., Sambrook, Fritsch & Maniatis, Molecular Cloning: A Laboratory Manual, Second Edition (1989) Cold Spring Harbor s Laboratory Press, Cold Spring Harbor, New York (herein "Sambrook et al., 1989") DNA Cloning: A Practical Approach, Volumes I and II /D.N. Glover ed. 1985); Oligonucleotide Synthesis (M.J. Gait ed. 1984); Nucleic Acid Hybridization (B.D. Hames & S.J. Higgins eds (1985)); Transcription And Translation (B.D. Hames & S.J. Higgins, eds. (1984)); Animal Cell Culture (R.I. Freshney, ed. (1986)); Immobilized Cells And Enzymes (IRL Press, 0 (1986)); B. Perbal, A Practical Guide To Molecular Cloning (1984).
When applied to a protein, the term "isolated" indicates that the protein is found in a condition other than its native environment, such as apart from blood and animal tissue. In a prefened form, the isolated protein is substantially free of other proteins, particularly other proteins of 5 animal origin. It is prefened to provide the proteins in a highly purified form, i.e., greater than 95% pure, more preferably greater than 99% pure. When applied to a polynucleotide molecule, the term "isolated" indicates that the molecule is removed from its natural genetic milieu, and is thus free of other extraneous or unwanted coding sequences, and is in a form suitable for use within genetically engineered protein production systems. Such isolated o molecules are those that are separated from their natural environment and include cDNA and genomic clones. Isolated DNA molecules of the present invention are free of other genes with
which they are ordinarily associated, and may include naturally occurring 5' and 3' untranslated regions such as promoters and terminators. The identification of associated regions will be evident to one of ordinary skill in the art (see for example, Dynan and Tijan, Nature 316: 774-78, 1985).
A "polynucleotide" is a single- or double-stranded polymer of deoxyribonucleotide or ribonucleotide bases read from the 5' to the 3' end. Polynucleotides include RNA and DNA, and may be isolated from natural sources, synthesized in vitro, or prepared from a combination of natural and synthetic molecules.
A "nucleic acid molecule" refers to the phosphate ester polymeric form of ribonucleosides (adenosine, guanosine, uridine or cytidine; "RNA molecules") or deoxyribonucleosides (deoxyadenosine, deoxyguanosine, deoxythymidine, or deoxycytidine; "DNA molecules") in either single stranded form, or a double-stranded helix. Double stranded DNA-DNA, DNA- RNA and RNA-RNA helices are possible. The term nucleic acid molecule, and in particular DNA or RNA molecule, refers only to the primary and secondary structure of the molecule, and does not limit it to any particular tertiary or quaternary forms. Thus, this term includes double-stranded DNA found, inter alia, in linear or circular DNA molecules (e.g., restriction fragments), plasmids, and chromosomes. In discussing the structure of particular double- stranded DNA molecules, sequences may be described herein according to the normal convention of giving only the sequence in the 5' to 3' direction along the nontranscribed strand of DNA (i.e., the strand having a sequence homologous to the mRNA). A "recombinant DNA molecule" is a DNA molecule that has undergone a molecular biological manipulation.
A DNA "coding sequence" is a double-stranded DNA sequence, which is transcribed and translated into a polypeptide in a cell in vitro or in vivo when placed under the control of appropriate regulatory sequences. The boundaries of the coding sequence are determined by a start codon at the 5' (amino) terminus and a translation stop codon at the 3' (carboxyl) terminus. A coding sequence can include, but is not limited to, prokaryotic sequences, cDNA from eukaryotic mRNA, genomic DNA sequences from eukaryotic (e.g., mammalian) DNA, and even synthetic DNA sequences. If the coding sequence is intended for expression in a
eukaryotic cell, a polyadenylation signal and transcription termination sequence will usually be located 3' to the coding sequence.
An "Expression vector" is a DNA molecule, linear or circular, that comprises a segment encoding a polypeptide of interest operably linked to additional segments that provide for its transcription. Such additional segments may include promoter and terminator sequences, and optionally one or more origins of replication, one or more selectable markers, an enhancer, a polyadenylation signal, and the like. Expression vectors are generally derived from plasmid or viral DNA, or may contain elements of both.
Transcriptional and translational control sequences are DNA regulatory sequences, such as promoters, enhancers, terminators, and the like, that provide for the expression of a coding sequence in a host cell. In eukaryotic cells, polyadenylation signals are control sequences.
A "secretory signal sequence" is a DNA sequence that encodes a polypeptide (a "secretory peptide" that, as a component of a larger polypeptide, directs the larger polypeptide through a secretory pathway of a cell in which it is synthesized. The larger polypeptide is commonly cleaved to remove the secretory peptide during transit through the secretory pathway.
The term "promoter" is used herein for its art-recognized meaning to denote a portion of a gene containing DNA sequences that provide for the binding of RNA polymerase and initiation of transcription. Promoter sequences are commonly, but not always, found in the 5' non-coding regions of genes.
"Operably linked", when referring to DNA segments, indicates that the segments are ananged so that they function in concert for their intended purposes, e.g. transcription initiates in the promoter and proceeds through the coding segment to the terminator.
A coding sequence is "under the control" of transcriptional and translational control sequences in a cell when RNA polymerase transcribes the coding sequence into mRNA, which is then trans-RNA spliced and translated into the protein encoded by the coding sequence.
"Isolated polypeptide" is a polypeptide which is essentially free of other non-[enzyme] polypeptides, e.g., at least about 20% pure, preferably at least about 40% pure, more preferably about 60% pure, even more preferably about 80% pure, most preferably about 90% pure, and even most preferably about 95% pure, as determined by SDS-PAGE.
"Heterologous" DNA refers to DNA not naturally located in the cell, or in a chromosomal site of the cell. Preferably, the heterologous DNA includes a gene foreign to the cell.
A cell has been "transfected" by exogenous or heterologous DNA when such DNA has been introduced inside the cell. A cell has been "transformed" by exogenous or heterologous DNA when the transfected DNA effects a phenotypic change. Preferably, the transforming DNA should be integrated (covalently linked) into chromosomal DNA making up the genome of the cell.
A "clone" is a population of cells derived from a single cell or common ancestor by mitosis.
"Homologous recombination" refers to the insertion of a foneign DNA sequence of a vector in a chromosome. Preferably, the vector targets a specific chromosomal site for homologous recombination. For specific homologous recombination, the vector will contain sufficiently long regions of homology to sequences of the chromosome to allow complementary binding and incorporation of the vector into the chromosome. Longer regions of homology, and greater degrees of sequence similarity, may increase the efficiency of homologous recombination.
Nucleic Acid Sequence
The techniques used to isolate or clone a nucleic acid sequence encoding a polypeptide are known in the art and include isolation from genomic DNA, preparation from cDNA, or a combination thereof. The cloning of the nucleic acid sequences of the present invention from such genomic DNA can be effected, e.g., by using the well known polymerase chain reaction (PCR) or antibody screening of expression libraries to detect cloned DNA fragments with shared structural features. See, e.g., Innis et al., 1990, A Guide to Methods and Application, Academic Press, New York. Other nucleic acid amplification procedures such as ligase chain
reaction (LCR), ligated activated transcription (LAT) and nuceic acid sequence-based amplification (NASBA) may be used. The nucleic acid sequence may be cloned from a strain producing the polypeptide, or from another related organism and thus, for example, may be an allelic or species variant of the polypeptide encoding region of the nucleic acid sequence.
The term "isolated" nucleic acid sequence as used herein refers to a nucleic acid sequence which is essentially free of other nucleic acid sequences, e.g., at least about 20% pure, preferably at least about 40% pure, more preferably about 60% pure, even more preferably about 80% pure, most preferably about 90% pure, and even most preferably about 95% pure, as determined by agarose gel electorphoresis. For example, an isolated nucleic acid sequence can be obtained by standard cloning procedures used in genetic engineering to relocate the nucleic acid sequence from its natural location to a different site where it will be reproduced. The cloning procedures may involve excision and isolation of a desired nucleic acid fragment comprising the nucleic acid sequence encoding the polypeptide, insertion of the fragment into a vector molecule, and incorporation of the recombinant vector into a host cell where multiple copies or clones of the nucleic acid sequence will be replicated. The nucleic acid sequence may be of genomic, cDNA, RNA, semisynthetic, synthetic origin, or any combinations thereof.
Nucleic Acid Construct
As used herein the term "nucleic acid construct" is intended to indicate any nucleic acid molecule of cDNA, genomic DNA, synthetic DNA or RNA origin. The term "construct" is intended to indicate a nucleic acid segment which may be single- or double-stranded, and which may be based on a complete or partial naturally occurring nucleotide sequence encoding a polypeptide of interest. The construct may optionally contain other nucleic acid segments.
The DNA of interest may suitably be of genomic or cDNA origin, for instance obtained by preparing a genomic or cDNA library and screening for DNA sequences coding for all or part of the polypeptide by hybridization using synthetic oligonucleotide probes in accordance with standard techniques (cf. Sambrook et al., supra).
The nucleic acid construct may also be prepared synthetically by established standard methods, e.g. the phosphoamidite method described by Beaucage and Caruthers, Tetrahedron Letters 22 (1981), 1859 - 1869, or the method described by Matthes et al, EMBO Journal 3 (1984), 801 - 805. According to the phosphoamidite method, oligonucleotides are synthesized, 5 e.g. in an automatic DNA synthesizer, purified, annealed, ligated and cloned in suitable vectors.
Furthermore, the nucleic acid construct may be of mixed synthetic and genomic, mixed synthetic and cDNA or mixed genomic and cDNA origin prepared by ligating fragments of o synthetic, genomic or cDNA origin (as appropriate), the fragments conesponding to various parts of the entire nucleic acid construct, in accordance with standard techniques.
The nucleic acid construct may also be prepared by polymerase chain reaction using specific primers, for instance as described in US 4,683,202 or Saiki et al., Science 239 (1988), 487 - s 491.
The term nucleic acid construct may be synonymous with the term expression cassette when the nucleic acid construct contains all the control sequences required for expression of a coding sequence of the present invention. The term "coding sequence" as defined herein is a o sequence which is transcribed into mRNA and translated into a polypeptide of the present invention when placed under the control of the above mentioned control sequences. The boundaries of the coding sequence are generally determined by a translation start codon ATG at the 5 '-terminus and a translation stop codon at the 3 '-terminus. A coding sequence can include, but is not limited to, DNA, cDNA, and recombinant nucleic acid sequences. 5
The term "control sequences" is defined herein to include all components which are necessary or advantageous for expression of the coding sequence of the nucleic acid sequence. Each control sequence may be native or foreign to the nucleic acid sequence encoding the polypeptide. Such control sequences include, but are not limited to, a leader, a 0 polyadenylation sequence, a propeptide sequence, a promoter, a signal sequence, and a transcription terminator. At a minimum, the control sequences include a promoter, and transcriptional and translational stop signals. The control sequences may be provided with
linkers for the purpose of introducing specific restriction sites facilitating ligation of the control sequences with the coding region of the nucleic acid sequence encoding a polypeptide.
The control sequence may be an appropriate promoter sequence, a nucleic acid sequence which is recognized by a host cell for expression of the nucleic acid sequence. The promoter sequence contains transcription and translation control sequences which mediate the expression of the polypeptide. The promoter may be any nucleic acid sequence which shows transcriptional activity in the host cell of choice and may be obtained from genes encoding extracellular or intracellular polypeptides either homologous or heterologous to the host cell. The control sequence may also be a suitable transcription terminator sequence, a sequence recognized by a host cell to terminate transcription. The terminator sequence is operably linked to the 3' terminus of the nucleic acid sequence encoding the polypeptide. Any terminator which is functional in the host cell of choice may be used in the present invention.
The control sequence may also be a polyadenylation sequence, a sequence which is operably linked to the 3' terminus of the nucleic acid sequence and which, when transcribed, is recognized by the host cell as a signal to add polyadenosine residues to transcribed mRNA. Any polyadenylation sequence which is functional in the host cell of choice may be used in the present invention.
The control sequence may also be a signal peptide coding region, which codes for an amino acid sequence linked to the amino terminus of the polypeptide which can direct the expressed polypeptide into the cell's secretory pathway of the host cell. The 5' end of the coding sequence of the nucleic acid sequence may inherently contain a signal peptide coding region naturally linked in translation reading frame with the segment of the coding region which encodes the secreted polypeptide. Alternatively, the 5' end of the coding sequence may contain a signal peptide coding region which is foreign to that portion of the coding sequence which encodes the secreted polypeptide. A foreign signal peptide coding region may be required where the coding sequence does not normally contain a signal peptide coding region.
Alternatively, the foreign signal peptide coding region may simply replace the natural signal peptide coding region in order to obtain enhanced secretion relative to the natural signal
peptide coding region normally associated with the coding sequence. The signal peptide coding region may be obtained from a glucoamylase or an amylase gene from an Aspergillus species, a lipase or proteinase gene from a Rhizomucor species, the gene for the alpha-factor from Saccharomyces cerevisiae, an amylase or a protease gene from a Bacillus species, or the calf preprochymosin gene. However, any signal peptide coding region capable of directing the expressed polypeptide into the secretory pathway of a host cell of choice may be used in the present invention.
The control sequence may also be a propeptide coding region, which codes for an amino acid sequence positioned at the amino terminus of a polypeptide. The resultant polypeptide is known as a proenzyme or propolypeptide (or a zymogen in some cases). A propolypeptide is generally inactive and can be converted to mature active polypeptide by catalytic or autocatalytic cleavage of the propeptide from the propolypeptide. The propeptide coding region may be obtained from the Bacillus subtilis alkaline protease gene (aprE), the Bacillus subtilis neutral protease gene (nprT), the Saccharomyces cerevisiae alpha-factor gene, or the Myceliophthora thermophilum laccase gene (WO 95/33836).
The nucleic acid constructs of the present invention may also comprise one or more nucleic acid sequences which encode one or more factors that are advantageous in the expression of the polypeptide, e.g., an activator (e.g., a trans-acting factor), a chaperone, and a processing protease. Any factor that is functional in the host cell of choice may be used in the present invention. The nucleic acids encoding one or more of these factors are not necessarily in tandem with the nucleic acid sequence encoding the polypeptide.
An activator is a protein which activates transcription of a nucleic acid sequence encoding a polypeptide (Kudla et al., 1990, EMBO Journal 9:1355-1364; Jarai and Buxton, 1994, Cunent Genetics 26:2238-244; Verdier, 1990, Yeast 6:271-297). The nucleic acid sequence encoding an activator may be obtained from the genes encoding Bacillus stearothermophilus NprA (nprA), Saccharomyces cerevisiae heme activator protein 1 (hapl), Saccharomyces cerevisiae galactose metabolizing protein 4 (gal4), and Aspergillus nidulans ammonia regulation protein (areA). For further examples, see Verdier, 1990, supra and MacKenzie et al, 1993, Journal of General Microbiology 139:2295-2307.
A chaperone is a protein which assists another polypeptide in folding properly (Haiti et al., 1994, TIBS 19:20-25; Bergeron et al, 1994, TIBS 19:124-128; Demolder et al., 1994, Journal of Biotechnology 32:179-189; Craig, 1993, Science 260:1902-1903; Gething and Sambrook, 1992, Nature 355:33-45; Puig and Gilbert, 1994, Journal of Biological Chemistry 269:7764- 5 7771; Wang and Tsou, 1993, The FASEB Journal 7:1515-11157; Robinson et al., 1994, Bio/Technology 1 :381-384). The nucleic acid sequence encoding a chaperone may be obtained from the genes encoding Bacillus subtilis GroE proteins, Aspergillus oryzae protein disulphide isomerase, Saccharomyces cerevisiae calnexin, Saccharomyces cerevisiae BΪP/GRP78, and Saccharomyces cerevisiae Hsp70. For further examples, see Gething and o Sambrook, 1992, supra, and Hartl et al., 1994, supra.
A processing protease is a protease that cleaves a propeptide to generate a mature biochemically active polypeptide (Enderlin and Ogrydziak, 1994, Yeast 10:67-79; Fuller et al., 1989, Proceedings of the National Academy of Sciences USA 86:1434-1438; Julius et al, s 1984, Cell 37:1075-1089; Julius et al., 1983, Cell 32:839-852). The nucleic acid sequence encoding a processing protease may be obtained from the genes encoding Aspergillus niger Kex2, Saccharomyces cerevisiae dipeptidylaminopeptidase, Saccharomyces cerevisiae Kex2, and Yanowia lipolytica dibasic processing endoprotease (xprό).
o It may also be desirable to add regulatory sequences which allow the regulation of the expression of the polypeptide relative to the growth of the host cell. Examples of regulatory systems are those which cause the expression of the gene to be turned on or off in response to a chemical or physical stimulus, including the presence of a regulatory compound. Regulatory systems in prokaryotic systems would include the lac, tac, and trp operator systems. In yeast, 5 the ADH2 system or GAL1 system may be used. In filamentous fungi, the TAKA alpha- amylase promoter, Aspergillus niger glucoamylase promoter, and the Aspergillus oryzae glucoamylase promoter may be used as regulatory sequences. Other examples of regulatory sequences are those which allow for gene amplification. In eukaryotic systems, these include the dihydrofolate reductase gene which is amplified in the presence of methotrexate, and the 0 metallothionein genes which are amplified with heavy metals. In these cases, the nucleic acid sequence encoding the polypeptide would be placed in tandem with the regulatory sequence.
Promoters
Examples of suitable promoters for directing the transcription of the nucleic acid constructs of the present invention, especially in a bacterial host cell, are the promoters obtained from the E. coli lac operon, the Streptomyces coelicolor agarase gene (dagA), the Bacillus subtilis levansucrase gene (sacB), the Bacillus subtilis alkaline protease gene, the Bacillus licheniformis alpha-amylase gene (amyL), the Bacillus stearothermophilus maltogenic amylase gene (amyM), the Bacillus amyloliquefaciens alpha-amylase gene (amyQ), the Bacillus amyloliquefaciens BAN amylase gene, the Bacillus licheniformis penicillinase gene (penP), the Bacillus subtilis xylA and xylB genes, and the prokaryotic beta-lactamase gene (Villa-Kamaroff et al., 1978, Proceedings of the National Academy of Sciences USA 75:3727- 3731), as well as the tac promoter (DeBoer et al., 1983, Proceedings of the National Academy of Sciences USA 80:21-25) , or the Bacillus pumilus xylosidase gene, or by the phage Lambda PR or PL promoters or the E. coli lac, trp or tac promoters. Further promoters are described in "Useful proteins from recombinant bacteria" in Scientific American, 1980, 242:74-94; and in Sambrook et al., 1989, supra.
Examples of suitable promoters for directing the transcription of the nucleic acid constructs of the present invention in a filamentous fungal host cell are promoters obtained from the genes encoding Aspergillus oryzae TAKA amylase, Rhizomucor miehei aspartic proteinase, Aspergillus niger neutral alpha-amylase, Aspergillus niger acid stable alpha-amylase, Aspergillus niger or Aspergillus awamori glucoamylase (glaA), Rhizomucor miehei lipase, Aspergillus oryzae alkaline protease, Aspergillus oryzae triose phosphate isomerase, Aspergillus nidulans acetamidase, Fusarium oxysporum trypsin-like protease (as described in U.S. Patent No. 4,288,627, which is incorporated herein by reference), and hybrids thereof. Particularly prefened promoters for use in filamentous fungal host cells are the TAKA amylase, NA2-tpi (a hybrid of the promoters from the genes encoding Aspergillus niger neutral (-amylase and Aspergillus oryzae triose phosphate isomerase), and glaA promoters. Further suitable promoters for use in filamentous fungus host cells are the ADH3 promoter (McKnight et al, The EMBO J. 4 (1985), 2093 - 2099) or the tpiA promoter.
Examples of suitable promoters for use in yeast host cells include promoters from yeast glycolytic genes (Hitzeman et al., J. Biol. Chem. 255 (1980), 12073 - 12080; Alber and
Kawasaki, J. Mol. Appl. Gen. 1 (1982), 419 - 434) or alcohol dehydrogenase genes (Young et al., in Genetic Engineering of Microorganisms for Chemicals (Hollaender et al, eds.), Plenum Press, New York, 1982), or the TPI1 (US 4,599,311) or ADH2-4c (Russell et al., Nature 304 (1983), 652 - 654) promoters.
Further useful promoters are obtained from the Saccharomyces cerevisiae enolase (ENO-1) gene, the Saccharomyces cerevisiae galactokinase gene (GAL1), the Saccharomyces cerevisiae alcohol dehydrogenase/glyceraldehyde-3-phosphate dehydrogenase genes (ADH2/GAP), and the Saccharomyces cerevisiae 3-phosphoglycerate kinase gene. Other useful promoters for yeast host cells are described by Romanos et al., 1992, Yeast 8:423-488. In a mammalian host cell, useful promoters include viral promoters such as those from Simian Virus 40 (SV40), Rous sarcoma virus (RSV), adenovirus, and bovine papilloma virus (BPV).
Examples of suitable promoters for directing the transcription of the DNA encoding the polypeptide of the invention in mammalian cells are the SV40 promoter (Subramani et al., Mol. Cell Biol. 1 (1981), 854 -864), the MT-1 (metallothionein gene) promoter (Palmiter et al., Science 222 (1983), 809 - 814) or the adenovirus 2 major late promoter. An example of a suitable promoter for use in insect cells is the polyhedrin promoter (US 4,745,051; Vasuvedan et al., FEBS Lett. 311, (1992) 7 - 11), the P10 promoter (J.M. Vlak et al., J. Gen. Virology 69, 1988, pp. 765-776), the Autographa califomica polyhedrosis virus basic protein promoter (EP 397 485), the baculovirus immediate early gene 1 promoter (US 5,155,037; US 5,162,222), or the baculovirus 39K delayed-early gene promoter (US 5,155,037; US 5,162,222).
Terminators
Prefened terminators for filamentous fungal host cells are obtained from the genes encoding Aspergillus oryzae TAKA amylase, Aspergillus niger glucoamylase, Aspergillus nidulans anthranilate synthase, Aspergillus niger alpha-glucosidase, and Fusarium oxysporum trypsin- like protease, for fungal hosts) the TPI1 (Alber and Kawasaki, op. cit.) or ADH3 (McKnight et al., op. cit.) terminators.
Prefened terminators for yeast host cells are obtained from the genes encoding Saccharomyces cerevisiae enolase, Saccharomyces cerevisiae cytochrome C (CYC1), or
Saccharomyces cerevisiae glyceraldehyde-3 -phosphate dehydrogenase. Other useful terminators for yeast host cells are described by Romanos et al, 1992, supra.
Polyadenylation Signals Prefened polyadenylation sequences for filamentous fungal host cells are obtained from the genes encoding Aspergillus oryzae TAKA amylase, Aspergillus niger glucoamylase, Aspergillus nidulans anthranilate synthase, and Aspergillus niger alpha-glucosidase. Useful polyadenylation sequences for yeast host cells are described by Guo and Sherman, 1995, Molecular Cellular Biology 15:5983-5990. Polyadenylation sequences are well known in the art for mammalian host cells such as SV40 or the adenovirus 5 Elb region.
Signal Sequences
An effective signal peptide coding region for bacterial host cells is the signal peptide coding region obtained from the maltogenic amylase gene from Bacillus NCEB 11837, the Bacillus stearothermophilus alpha-amylase gene, the Bacillus licheniformis subtilisin gene, the Bacillus licheniformis beta-lactamase gene, the Bacillus stearothermophilus neutral proteases genes (nprT, nprS, nprM), and the Bacillus subtilis PrsA gene. Further signal peptides are described by Simonen and Palva, 1993, Microbiological Reviews 57:109-137.
An effective signal peptide coding region for filamentous fungal host cells is the signal peptide coding region obtained from Aspergillus oryzae TAKA amylase gene, Aspergillus niger neutral amylase gene, the Rhizomucor miehei aspartic proteinase gene, the Humicola lanuginosa cellulase or lipase gene, or the Rhizomucor miehei lipase or protease gene, Aspergillus sp. amylase or glucoamylase, a gene encoding a Rhizomucor miehei lipase or protease. The signal peptide is preferably derived from a gene encoding A. oryzae TAKA amylase, A. niger neutral (-amylase, A. niger acid-stable amylase, or A. niger glucoamylase.
Useful signal peptides for yeast host cells are obtained from the genes for Saccharomyces cerevisiae a- factor and Saccharomyces cerevisiae invertase. Other useful signal peptide coding regions are described by Romanos et al., 1992, supra.
For secretion from yeast cells, the secretory signal sequence may encode any signal peptide which ensures efficient direction of the expressed polypeptide into the secretory pathway of the cell. The signal peptide may be naturally occurring signal peptide, or a functional part thereof, or it may be a synthetic peptide. Suitable signal peptides have been found to be the a- factor signal peptide (cf. US 4,870,008), the signal peptide of mouse salivary amylase (cf. O. Hagenbuchle et al., Nature 289, 1981, pp. 643-646), a modified carboxypeptidase signal peptide (cf. L.A. Vails et al., Cell 48, 1987, pp. 887-897), the yeast BAR1 signal peptide (cf. WO 87/02670), or the yeast aspartic protease 3 (YAP3) signal peptide (cf. M. Egel-Mitani et al, Yeast 6, 1990, pp. 127-137). For efficient secretion in yeast, a sequence encoding a leader peptide may also be inserted downstream of the signal sequence and uptream of the DNA sequence encoding the polypeptide. The function of the leader peptide is to allow the expressed polypeptide to be directed from the endoplasmic reticulum to the Golgi apparatus and further to a secretory vesicle for secretion into the culture medium (i.e. exportation of the polypeptide across the cell wall or at least through the cellular membrane into the periplasmic space of the yeast cell). The leader peptide may be the yeast a- factor leader (the use of which is described in e.g. US 4,546,082, EP 16 201, EP 123 294, EP 123 544 and EP 163 529). Alternatively, the leader peptide may be a synthetic leader peptide, which is to say a leader peptide not found in nature. Synthetic leader peptides may, for instance, be constructed as described in WO 89/02463 or WO 92/11378.
For use in insect cells, the signal peptide may conveniently be derived from an insect gene (cf. WO 90/05783), such as the lepidopteran Manduca sexta adipokinetic hormone precursor signal peptide (cf. US 5,023,328).
Expression Vectors
The present invention also relates to recombinant expression vectors comprising a nucleic acid sequence of the present invention, a promoter, and transcriptional and translational stop signals. The various nucleic acid and control sequences described above may be joined together to produce a recombinant expression vector which may include one or more convenient restriction sites to allow for insertion or substitution of the nucleic acid sequence encoding the polypeptide at such sites. Alternatively, the nucleic acid sequence of the present
invention may be expressed by inserting the nucleic acid sequence or a nucleic acid construct comprising the sequence into an appropriate vector for expression. In creating the expression vector, the coding sequence is located in the vector so that the coding sequence is operably linked with the appropriate control sequences for expression, and possibly secretion.
The recombinant expression vector may be any vector (e.g., a plasmid or virus) which can be conveniently subjected to recombinant DNA procedures and can bring about the expression of the nucleic acid sequence. The choice of the vector will typically depend on the compatibility of the vector with the host cell into which the vector is to be introduced. The vectors may be linear or closed circular plasmids. The vector may be an autonomously replicating vector, i.e., a vector which exists as an extrachromosomal entity, the replication of which is independent of chromosomal replication, e.g., a plasmid, an extrachromosomal element, a minichromosome, or an artificial chromosome. The vector may contain any means for assuring self-replication. Alternatively, the vector may be one which, when introduced into the host cell, is integrated into the genome and replicated together with the chromosome(s) into which it has been integrated. The vector system may be a single vector or plasmid or two or more vectors or plasmids which together contain the total DNA to be introduced into the genome of the host cell, or a transposon.
The vectors of the present invention preferably contain one or more selectable markers which permit easy selection of transformed cells. A selectable marker is a gene the product of which provides for biocide or viral resistance, resistance to heavy metals, prototrophy to auxotrophs, and the like. Examples of bacterial selectable markers are the dal genes from Bacillus subtilis or Bacillus licheniformis, or markers which confer antibiotic resistance such as ampicillin, kanamycin, chloramphenicol, tetracycline, neomycin, hygromycin or methotrexate resistance. A frequently used mammalian marker is the dihydrofolate reductase gene (DHFR). Suitable markers for yeast host cells are ADE2, HIS3, LEU2, LYS2, MET3, TRP1, and URA3. A selectable marker for use in a filamentous fungal host cell may be selected from the group including, but not limited to, amdS (acetamidase), argB (ornithine carbamoyltransferase), bar (phosphinothricin acetyltransferase), hygB (hygromycin phosphotransferase), niaD (nitrate reductase), pyrG (orotidine-5' -phosphate decarboxylase), sC (sulfate adenyltransferase), trpC (anthranilate synthase), and glufosinate resistance markers, as well as equivalents from other
species. Prefened for use in an Aspergillus cell are the amdS and pyrG markers of Aspergillus nidulans or Aspergillus oryzae and the bar marker of Streptomyces hygroscopicus. Furthermore, selection may be accomplished by co-transformation, e.g., as described in WO 91/17243, where the selectable marker is on a separate vector.
The vectors of the present invention preferably contain an element(s) that permits stable integration of the vector into the host cell genome or autonomous replication of the vector in the cell independent of the genome of the cell.
The vectors of the present invention may be integrated into the host cell genome when introduced into a host cell. For integration, the vector may rely on the nucleic acid sequence encoding the polypeptide or any other element of the vector for stable integration of the vector into the genome by homologous or nonhomologous recombination. Alternatively, the vector may contain additional nucleic acid sequences for directing integration by homologous recombination into the genome of the host cell. The additional nucleic acid sequences enable the vector to be integrated into the host cell genome at a precise location(s) in the chromosome(s). To increase the likelihood of integration at a precise location, the integrational elements should preferably contain a sufficient number of nucleic acids, such as 100 to 1,500 base pairs, preferably 400 to 1,500 base pairs, and most preferably 800 to 1,500 base pairs, which are highly homologous with the conesponding target sequence to enhance the probability of homologous recombination. The integrational elements may be any sequence that is homologous with the target sequence in the genome of the host cell. Furthermore, the integrational elements may be non-encoding or encoding nucleic acid sequences. On the other hand, the vector may be integrated into the genome of the host cell by non-homologous recombination. These nucleic acid sequences may be any sequence that is homologous with a target sequence in the genome of the host cell, and, furthermore, may be non-encoding or encoding sequences.
For autonomous replication, the vector may further comprise an origin of replication enabling the vector to replicate autonomously in the host cell in question. Examples of bacterial origins of replication are the origins of replication of plasmids pBR322, pUC19, pACYC177, pACYC184, pUBHO, pE194, pTA1060, and pAMBl. Examples of origin of replications for
use in a yeast host cell are the 2 micron origin of replication, the combination of CEN6 and ARS4, and the combination of CEN3 and ARS1. The origin of replication may be one having a mutation which makes its functioning temperature-sensitive in the host cell (see, e.g., Ehrlich, 1978, Proceedings of the National Academy of Sciences USA 75:1433).
More than one copy of a nucleic acid sequence encoding a polypeptide of the present invention may be inserted into the host cell to amplify expression of the nucleic acid sequence. Stable amplification of the nucleic acid sequence can be obtained by integrating at least one additional copy of the sequence into the host cell genome using methods well known in the art and selecting for transformants.
The procedures used to ligate the elements described above to construct the recombinant expression vectors of the present invention are well known to one skilled in the art (see, e.g., Sambrook et al., 1989, supra).
Host Cells
The present invention also relates to recombinant host cells, comprising a nucleic acid sequence of the invention, which are advantageously used in the recombinant production of the polypeptides. The term "host cell" encompasses any progeny of a parent cell which is not identical to the parent cell due to mutations that occur during replication.
The cell is preferably transformed with a vector comprising a nucleic acid sequence of the invention followed by integration of the vector into the host chromosome. "Transformation" means introducing a vector comprising a nucleic acid sequence of the present invention into a host cell so that the vector is maintained as a chromosomal integrant or as a self-replicating extra-chromosomal vector. Integration is generally considered to be an advantage as the nucleic acid sequence is more likely to be stably maintained in the cell. Integration of the vector into the host chromosome may occur by homologous or non-homologous recombination as described above.
The choice of a host cell will to a large extent depend upon the gene encoding the polypeptide and its source. The host cell may be a unicellular microorganism, e.g., a prokaryote, or a non- unicellular microorganism, e.g., a eukaryote. Useful unicellular cells are bacterial cells such
as gram positive bacteria including, but not limited to, a Bacillus cell, e.g., Bacillus alkalophilus, Bacillus amyloliquefaciens, Bacillus brevis, Bacillus circulans, Bacillus coagulans, Bacillus lautus, Bacillus lentus, Bacillus licheniformis, Bacillus megaterium, Bacillus stearothermophilus, Bacillus subtilis, and Bacillus thuringiensis; or a Streptomyces cell, e.g., Streptomyces lividans or Streptomyces murinus, or gram negative bacteria such as E. coli and Pseudomonas sp. In a prefened embodiment, the bacterial host cell is a Bacillus lentus, Bacillus licheniformis, Bacillus stearothermophilus or Bacillus subtilis cell. The transformation of a bacterial host cell may, for instance, be effected by protoplast transformation (see, e.g., Chang and Cohen, 1979, Molecular General Genetics 168:111-115), by using competent cells (see, e.g., Young and Spizizin, 1961, Journal of Bacteriology 81 :823-829, or Dubnar and Davidoff-Abelson, 1971, Journal of Molecular Biology 56:209- 221), by electroporation (see, e.g., Shigekawa and Dower, 1988, Biotechniques 6:742-751), or by conjugation (see, e.g., Koehler and Thome, 1987, Journal of Bacteriology 169:5771-5278).
The host cell may be a eukaryote, such as a mammalian cell, an insect cell, a plant cell or a fungal cell.
Useful mammalian cells include Chinese hamster ovary (CHO) cells, HeLa cells, baby hamster kidney (BHK) cells, COS cells, or any number of other immortalized cell lines available, e.g., from the American Type Culture Collection.
Examples of suitable mammalian cell lines are the COS (ATCC CRL 1650 and 1651), BHK (ATCC CRL 1632, 10314 and 1573, ATCC CCL 10), CHL (ATCC CCL39) or CHO (ATCC CCL 61) cell lines. Methods of transfecting mammalian cells and expressing DNA sequences introduced in the cells are described in e.g. Kaufman and Sharp, J. Mol. Biol. 159 (1982), 601 - 621; Southern and Berg, J. Mol. Appl. Genet. 1 (1982), 327 - 341; Loyter et al, Proc. Natl. Acad. Sci. USA 79 (1982), 422 - 426; Wigler et al, Cell 14 (1978), 725; Corsaro and Pearson, Somatic Cell Genetics 7 (1981), 603, Ausubel et al., Cunent Protocols in Molecular Biology, John Wiley and Sons, Inc., N.Y., 1987, Hawley-Nelson et al., Focus 15 (1993), 73; Ciccarone et al., Focus 15 (1993), 80; Graham and van der Eb, Virology 52 (1973), 456; and Neumann et al., EMBO J. 1 (1982), 841 - 845.
In a prefened embodiment, the host cell is a fungal cell. "Fungi" as used herein includes the phyla Ascomycota, Basidiomycota, Chytridiomycota, and Zygomycota (as defined by Hawksworth et al., In, Ainsworth and Bisby's Dictionary of The Fungi, 8th edition, 1995, CAB International, University Press, Cambridge, UK) as well as the Oomycota (as cited in Hawksworth et al., 1995, supra, page 171) and all mitosporic fungi (Hawksworth et al., 1995, supra). Representative groups of Ascomycota include, e.g., Neurospora, Eupenicillium (=Penicillium), Emericella (=Aspergillus), Eurotium (=Aspergillus), and the true yeasts listed above. Examples of Basidiomycota include mushrooms, rusts, and smuts. Representative groups of Chytridiomycota include, e.g., Allomyces, Blastocladiella, Coelomomyces, and aquatic fungi. Representative groups of Oomycota include, e.g., Saprolegniomycetous aquatic fungi (water molds) such as Achlya. Examples of mitosporic fungi include Aspergillus, Penicillium, Candida, and Alternaria. Representative groups of Zygomycota include, e.g., Rhizopus and Mucor. In a prefened embodiment, the fungal host cell is a yeast cell. "Yeast" as used herein includes ascosporogenous yeast (Endomycetales), basidiosporogenous yeast, and yeast belonging to the Fungi Imperfecti (Blastomycetes). The ascosporogenous yeasts are divided into the families Spermophthoraceae and Saccharomycetaceae. The latter is comprised of four subfamilies, Schizosaccharomycoideae (e.g., genus Schizosaccharomyces), Nadsonioideae, Lipomycoideae, and Saccharomycoideae (e.g., genera Pichia, Kluyveromyces and Saccharomyces). The basidiosporogenous yeasts include the genera Leucosporidim, Rhodosporidium, Sporidiobolus, Filobasidium, and Filobasidiella. Yeast belonging to the Fungi Imperfecti are divided into two families, Sporobolomycetaceae (e.g., genera Sorobolomyces and Bullera) and Cryptococcaceae (e.g., genus Candida). Since the classification of yeast may change in the future, for the purposes of this invention, yeast shall be defined as described in Biology and Activities of Yeast (Skinner, F.A., Passmore, S.M., and Davenport, R.R., eds, Soc. App. Bacteriol. Symposium Series No. 9, 1980. The biology of yeast and manipulation of yeast genetics are well known in the art (see, e.g., Biochemistry and Genetics of Yeast, Bacil, M., Horecker, B.J., and Stopani, A.O.M., editors, 2nd edition, 1987; The Yeasts, Rose, A.H., and Harrison, J.S., editors, 2nd edition, 1987; and The Molecular Biology of the Yeast Saccharomyces, Strathern et al., editors, 1981).
The yeast host cell may be selected from a cell of a species of Candida, Kluyveromyces, Saccharomyces, Schizosaccharomyces, Candida, Pichia, Hansehula, , or Yanowia. In a prefened embodiment, the yeast host cell is a Saccharomyces carlsbergensis, Saccharomyces cerevisiae, Saccharomyces diastaticus, Saccharomyces douglasii, Saccharomyces kluyveri, Saccharomyces norbensis or Saccharomyces oviformis cell. Other useful yeast host cells are a Kluyveromyces lactis Kluyveromyces fragilis Hansehula polymorpha, Pichia pastoris Yanowia lipolytica, Schizosaccharomyces pombe, Ustilgo maylis, Candida maltose, Pichia guillermondii and Pichia methanolio cell (cf. Gleeson et al., J. Gen. Microbiol. 132, 1986, pp. 3459-3465; US 4,882,279 and US 4,879,231).
In a prefened embodiment, the fungal host cell is a filamentous fungal cell. "Filamentous fungi" include all filamentous forms of the subdivision Eumycota and Oomycota (as defined by Hawksworth et al, 1995, supra). The filamentous fungi are characterized by a vegetative mycelium composed of chitin, cellulose, glucan, chitosan, mannan, and other complex polysaccharides. Vegetative growth is by hyphal elongation and carbon catabolism is obligately aerobic. In contrast, vegetative growth by yeasts such as Saccharomyces cerevisiae is by budding of a unicellular thallus and carbon catabolism may be fermentative. In a more prefened embodiment, the filamentous fungal host cell is a cell of a species of, but not limited to, Acremonium, Aspergillus, Fusarium, Humicola, Mucor, Myceliophthora, Neurospora, Penicillium, Thielavia, Tolypocladium, and Trichoderma or a teleomorph or synonym thereof. In an even more prefened embodiment, the filamentous fungal host cell is an Aspergillus cell. In another even more prefened embodiment, the filamentous fungal host cell is an Acremonium cell. In another even more prefened embodiment, the filamentous fungal host cell is a Fusarium cell. In another even more prefened embodiment, the filamentous fungal host cell is a Humicola cell. In another even more prefened embodiment, the filamentous fungal host cell is a Mucor cell. In another even more prefened embodiment, the filamentous fungal host cell is a Myceliophthora cell. In another even more prefened embodiment, the filamentous fungal host cell is a Neurospora cell. In another even more prefened embodiment, the filamentous fungal host cell is a Penicillium cell. In another even more prefened embodiment, the filamentous fungal host cell is a Thielavia cell. In another even more prefened embodiment, the filamentous fungal host cell is a Tolypocladium cell. In another even more prefened embodiment, the filamentous fungal host cell is a Trichoderma
cell. In a most prefened embodiment, the filamentous fungal host cell is an Aspergillus awamori, Aspergillus foetidus, Aspergillus japonicus, Aspergillus niger, Aspergillus nidulans or Aspergillus oryzae cell. In another most prefened embodiment, the filamentous fungal host cell is a Fusarium cell of the section Discolor (also known as the section Fusarium). For example, the filamentous fungal parent cell may be a Fusarium bactridioides, Fusarium cerealis, Fusarium crookwellense, Fusarium culmorum, Fusarium graminearum, Fusarium graminum, Fusarium heterosporum, Fusarium negundi, Fusarium reticulatum, Fusarium roseum, Fusarium sambucinum, Fusarium sarcochroum, Fusarium sulphureum, or Fusarium trichothecioides cell. In another prefered embodiment, the filamentous fungal parent cell is a Fusarium strain of the section Elegans, e.g., Fusarium oxysporum. In another most prefened embodiment, the filamentous fungal host cell is a Humicola insolens or Humicola lanuginosa cell. In another most prefened embodiment, the filamentous fungal host cell is a Mucor miehei cell. In another most prefened embodiment, the filamentous fungal host cell is a Myceliophthora thermophilum cell. In another most prefened embodiment, the filamentous fungal host cell is a Neurospora crassa cell. In another most prefened embodiment, the filamentous fungal host cell is a Penicillium purpurogenum cell. In another most prefened embodiment, the filamentous fungal host cell is a Thielavia tenestris cell or a Acremonium chrysogenum cell. In another most prefened embodiment, the Trichoderma cell is a Trichoderma harzianum, Trichoderma koningii, Trichoderma longibrachiatum, Trichoderma reesei or Trichoderma viride cell. The use of Aspergillus spp. for the expression of proteins is described in, e.g., EP 272 277, EP 230 023.
Transformation
Fungal cells may be transformed by a process involving protoplast formation, transformation of the protoplasts, and regeneration of the cell wall in a manner known per se. Suitable procedures for transformation of Aspergillus host cells are described in EP 238 023 and Yelton et al, 1984, Proceedings of the National Academy of Sciences USA 81:1470-1474. A suitable method of transforming Fusarium species is described by Malardier et al., 1989, Gene 78:147-156 or in copending US Serial No. 08/269,449. Examples of other fungal cells are cells of filamentous fungi, e.g. Aspergillus spp., Neurospora spp., Fusarium spp. or
Trichoderma spp., in particular strains of A. oryzae, A. nidulans or A. niger. The use of Aspergillus spp. for the expression of proteins is described in, e.g., EP 272 277, EP 230 023,
EP 184 ... The transformation of F. oxysporum may, for instance, be carried out as described by Malardier et al., 1989, Gene 78: 147-156.
Yeast may be transformed using the procedures described by Becker and Guarente, In Abelson, J.N. and Simon, M.I., editors, Guide to Yeast Genetics and Molecular Biology, Methods in Enzymology, Volume 194, pp 182-187, Academic Press, Inc., New York; Ito et al., 1983, Journal of Bacteriology 153:163; and Hinnen et al., 1978, Proceedings of the National Academy of Sciences USA 75:1920. Mammalian cells may be transformed by direct uptake using the calcium phosphate precipitation method of Graham and Van der Eb (1978, Virology 52:546).
Transformation of insect cells and production of heterologous polypeptides therein may be performed as described in US 4,745,051; US 4, 775, 624; US 4,879,236; US 5,155,037; US 5,162,222; EP 397,485) all of which are incoφorated herein by reference. The insect cell line used as the host may suitably be a Lepidoptera cell line, such as Spodoptera frugiperda cells or Trichoplusia ni cells (cf. US 5,077,214). Culture conditions may suitably be as described in, for instance, WO 89/01029 or WO 89/01028, or any of the aforementioned references.
Methods of Production The transformed or transfected host cells described above are cultured in a suitable nutrient medium under conditions permitting the production of the desired molecules, after which these are recovered from the cells, or the culture broth.
The medium used to culture the cells may be any conventional medium suitable for growing the host cells, such as minimal or complex media containing appropriate supplements. Suitable media are available from commercial suppliers or may be prepared according to published recipes (e.g. in catalogues of the American Type Culture Collection). The media are prepared using procedures known in the art (see, e.g., references for bacteria and yeast; Bennett, J.W. and LaSure, L., editors, More Gene Manipulations in Fungi, Academic Press, CA, 1991).
If the molecules are secreted into the nutrient medium, they can be recovered directly from the medium. If they are not secreted, they can be recovered from cell lysates. The molecules are recovered from the culture medium by conventional procedures including separating the host cells from the medium by centrifugation or filtration, precipitating the proteinaceous components of the supernatant or filtrate by means of a salt, e.g. ammonium sulphate, purification by a variety of chromatographic procedures, e.g. ion exchange chromatography, gelfiltration chromatography, affinity chromatography, or the like, dependent on the type of molecule in question.
The molecules of interest may be detected using methods known in the art that are specific for the molecules. These detection methods may include use of specific antibodies, formation of a product, or disappearance of a substrate. For example, an enzyme assay may be used to determine the activity of the molecule. Procedures for determining various kinds of activity are known in the art.
The molecules of the present invention may be purified by a variety of procedures known in the art including, but not limited to, chromatography (e.g., ion exchange, affinity, hydrophobic, chromatofocusing, and size exclusion), electrophoretic procedures (e.g., preparative isoelectric focusing (IEF), differential solubility (e.g., ammonium sulfate precipitation), or extraction (see, e.g., Protein Purification, J-C Janson and Lars Ryden, editors, VCH Publishers, New York, 1989).
The term "immunological response", used in connection with the present invention, is the response of an organism to a compound, which involves the immune system according to any of the four standard reactions (Type I, II, III and IV according to Coombs & Gell).
Conespondingly, the "immunogenicity" of a compound used in connection with the present invention refers to the ability of this compound to induce an 'immunological response' in animals including man.
The term "allergic response", used in connection with the present invention, is the response of an organism to a compound, which involves IgE mediated responses (Type I reaction
according to Coombs & Gell). It is to be understood that sensibilization (i.e. development of compound-specific IgE antibodies) upon exposure to the compound is included in the definition of "allergic response".
Conespondingly, the "allergenicity" of a compound used in connection with the present 5 invention refers to the ability of this compound to induce an 'allergic response' in animals including man.
The terms "relevant protein backbone" or 'protein backbone' refer to the polypeptide to be modified by creating a library of diversified mutants. The "relevant protein backbone" may be o a naturally occurring (or wild-type) polypeptide or it may be a variant thereof prepared by any suitable means. For instance, the "relevant protein backbone" may be a variant of a naturally occurring polypeptide which has been modified by substitution, deletion or truncation of one or more amino acid residues or by addition or insertion of one or more amino acid residues to the amino acid sequence of a naturally-occurring polypeptide. 5
The term " randomized library" of protein variants refers to a library with at least partially randomized composition of the members, e.g. protein variants.
The term "functionality" of protein variants refers to e.g. enzymatic activity, binding to a 0 ligand or receptor, stimulation of a cellular response (e.g. 3H-thymidine incorporation as response to a mitogenic factor), or anti-microbial activity.
An "epitope" is a set of amino acids on a protein that are involved in an immunological response, such as antibody binding or T-cell activation. One particularly useful method of 5 identifying epitopes involved in antibody binding is to screen a library of peptide-phage membrane protein fusions and selecting those that bind to relevant antigen-specific antibodies, sequencing the randomized part of the fusion gene, aligning the sequences involved in binding, defining consensus sequences based on these alignments, and mapping these consensus sequences on the surface or the sequence and/or structure of the antigen, to identify o epitopes involved in antibody binding.
By the term "epitope pattern" is meant such a consensus sequence of peptides that bind well to a relevant antibody.
An "epitope area" is defined as the amino acids situated within 5 A from the epitope amino acids. Modifications of amino acids of the 'epitope area' can possibly affect the function of the conesponding epitope.
By the term "specific polyclonal antibodies" is meant polyclonal antibodies isolated according to their specificity for a certain antigen, e.g. the protein backbone.
By the term "monospecific antibodies" is meant polyclonal antibodies isolated according to their specificity for a certain 'epitope pattern'. Such monospecific antibodies will only bind to one epitope pattern, but they may very well be produced by a number of antibody producing cells and recognize the same epitope patterns, thereby being polyclonal.
'Spiked mutagenesis' is a form of site-directed mutagenesis, in which the primers used have been synthesized using mixtures of oligonucleotides at one or more positions.
Detailed description of the invention
The inventors have found a method for high throughput screening (HTS) of a large population of host cells for production of a molecule of interest.
By applying the present invention to a diversified library of protein variants it is possible to screen a large number of protein variants for their ability to bind to specific antibodies in a quick and automated manner thereby providing leads that may be tested for their immunogenicity in animal studies.
The present invention relates to method for screening a library of protein variants for functional variants with reduced antibody binding capacity, comprising the steps of:
(i) generating a diversified library of protein variants starting from a relevant protein backbone,
(ii) transforming the library into suitable host cells,
(iii) culturing host cells,
(iv) sampling each cell culture ,
(v) analysing a sample by determining the antibody binding capacity of the variant protein,
(vi) analysing a sample by determining the functionality of the variant protein.
Protein backbone
The "relevant protein backbone" can in principle be any protein molecule of biological origin, non-limiting examples of which are peptides, polypeptides, proteins, enzymes, post- translationally modified polypeptides such as lipopeptides or glycosylated peptides, antimicrobial peptides or molecules, and proteins having pharmaceutical properties etc.
Accordingly the invention relates to a method, wherein the "relevant protein backbone" is chosen from the group consisting of polypeptides, small peptides, lipopeptides, antimicrobials, and pharmaceutical polypeptides.
The term "pharmaceutical polypeptides" is defined as polypeptides, including peptides, such as peptide hormones, proteins and/or enzymes, being physiologically active when administered to humans and/or animals.
Examples of "pharmaceutical polypeptides" contemplated according to the invention include insulin, ACTH, glucagon, somatostatin, somatotropin, thymosin, parathyroid hormone, pigmentary hormones, somatomedin, erythropoietin, luteinizing hormone, chorionic gonadotropin, hypothalmic releasing factors, antidiuretic hormones, thyroid stimulating
hormone, relaxin, interferons, thrombopoietin (TPO), blood coagulation factors, plasminogen activators such as streptokinase and tissue-plasminogen activator, cerebrosidase, and prolactin.
However, the proteins are preferably to be used in industry, housekeeping and/or medicine, such as proteins used in personal care products (for example shampoo; soap; skin, hand and face lotions; skin, hand and face creams; hair dyes; toothpaste), food/feed (for example in the baking industry), detergents, anti-microbial compositions.
In one embodiment of the invention the protein is an enzyme, such as glycosyl hydrolases, carbohydrases, peroxidases, proteases, lipases, phytases, polysaccharide lyases, oxidoreductases, transglutaminases and glycoseisomerases, in particular the following.
Parent Proteases
Parent proteases (i.e. enzymes classified under the Enzyme Classification number E.C. 3.4 in accordance with the Recommendations (1992) of the International Union of Biochemistry and
Molecular Biology (IUBMB)) include proteases within this group.
Examples include proteases selected from those classified under the Enzyme Classification
(E.C.) numbers:
3.4.11 (i.e. so-called aminopeptidases), including 3.4.11.5 (Prolyl aminopeptidase), 3.4.11.9 (X-pro aminopeptidase), 3.4.11.10 (Bacterial leucyl aminopeptidase), 3.4.11.12 (Thermophihc aminopeptidase), 3.4.11.15 (Lysyl aminopeptidase), 3.4.11.17 (Tryptophanyl aminopeptidase),
3.4.11.18 (Methionyl aminopeptidase).
3.4.21 (i.e. so-called serine endopeptidases), including 3.4.21.1 (Chymotrypsin), 3.4.21.4 (Trypsin), 3.4.21.25 (Cucumisin), 3.4.21.32 (Brachyurin), 3.4.21.48 (Cerevisin) and 3.4.21.62 (Subtilisin);
3.4.22 (i.e. so-called cysteine endopeptidases), including 3.4.22.2 (Papain), 3.4.22.3 (Ficain), 3.4.22.6 (Chymopapain), 3.4.22.7 (Asclepain), 3.4.22.14 (Actinidain), 3.4.22.30 (Caricain) and 3.4.22.31 (Ananain);
3.4.23 (i.e. so-called aspartic endopeptidases), including 3.4.23.1 (Pepsin A), 3.4.23.18 (Aspergillopepsin I), 3.4.23.20 (Penicillopepsin) and 3.4.23.25 (Saccharopepsin); and
3.4.24 (i.e. so-called metalloendopeptidases), including 3.4.24.28 (Bacillolysin).
Examples of relevant subtilisins comprise subtilisin BPN', subtilisin amylosacchariticus, subtilisin 168, subtilisin mesentericopeptidase, subtilisin Carlsberg, subtilisin DY, subtilisin 309, subtilisin 147, theπnitase, aqualysin, Bacillus PB92 protease, proteinase K, Protease TW7, and Protease TW3. 5 Specific examples of such readily available commercial proteases include Esperase®, Alcalase®, Neutrase®, Dyrazym®, Savinase®, Pyrase®, Pancreatic Trypsin NOVO (PTN), Bio-Feed( Pro, Clear-Lens Pro (all enzymes available from Novo Nordisk A/S). Examples of other commercial proteases include Maxtase®, Maxacal®, Maxapem® marketed by Gist-Brocades N.V., Opticlean® marketed by Solvay et Cie. and Purafect® marketed by l o Genencor International .
It is to be understood that also protease variants are contemplated as the parent protease. Examples of such protease variants are disclosed in EP 130.756 (Genentech), EP 214.435 (Henkel), WO 87/04461 (Amgen), WO 87/05050 (Genex), EP 251.446 (Genencor), EP 260.105 (Genencor), Thomas et al., (1985), Nature. 318, p. 375-376, Thomas et al., (1987), J. is Mol. Biol., 193, pp. 803-813, Russel et al., (1987), Nature, 328, p. 496-500, WO 88/08028 (Genex), WO 88/08033 (Amgen), WO 89/06279 (Novo Nordisk A/S), WO 91/00345 (Novo Nordisk A/S), EP 525 610 (Solvay) and WO 94/02618 (Gist-Brocades N.V.). The activity of proteases can be determined as described in "Methods of Enzymatic Analysis", third edition, 1984, Verlag Chemie, Weinheim, vol. 5.
20
Parent Lipases
Parent lipases (i.e. enzymes classified under the Enzyme Classification number E.C. 3.1.1
(Carboxylic Ester Hydrolases) in accordance with the Recommendations (1992) of the
International Union of Biochemistry and Molecular Biology (IUBMB)) include lipases within 5 this group.
Examples include lipases selected from those classified under the Enzyme Classification
(E.C.) numbers:
3.1.1 (i.e. so-called Carboxylic Ester Hydrolases), including (3.1.1.3) Triacylglycerol lipases,
(3.1.1.4.) Phosphorlipase A2. 0 Examples of lipases include lipases derived from the following microorganisms. The indicated patent publications are incorporated herein by reference:
Humicola, e.g. H. brevispora, H. lanuginosa, H. brevis var. thermoidea and H. insolens (US
4,810,414).
Pseudomonas, e.g. Ps. fragi, Ps. stutzeri, Ps. cepacia and Ps. fluorescens (WO 89/04361), or
Ps. plantarii or Ps. gladioli (US patent no. 4,950,417 (Solvay enzymes)) or Ps. alcaligenes and 5 Ps. pseudoalcaligenes (EP 218 272) or Ps. mendocina (WO 88/09367; US 5,389,536).
Fusarium, e.g. F. oxyspomm (EP 130,064) or F. solani pisi(WO 90/09446). Mucor (also called Rhizomucor), e.g. M. miehei (EP 238 023). Chromobacterium
(especially C. viscosum). Aspergillus (especially A. niger). Candida, e.g. C. cylindracea (also called C. rugosa) or C. antarctica (WO 88/02775) or C. antarctica lipase A or B (WO o 94/01541 and WO 89/02916). Geotricum, e.g. G. candidum (Schimada et al., (1989), J.
Biochem., 106, 383-388). Penicillium, e.g. P. camembertii (Yamaguchi et al., (1991), Gene
103, 61-67). Rhizopus, e.g. R. delemar (Hass et al., (1991), Gene 109, 107-113) or R. niveus (Kugimiya et al., (1992) Biosci. Biotech. Biochem 56, 716-719) or R. oryzae.
Bacillus, e.g. B. subtilis (Dartois et al, (1993) s Biochemica et Biophysica acta 1131, 253-260) or B. stearothermophilus (JP 64/7744992) or
B. pumilus (WO 91/16422).
Specific examples of readily available commercial lipases include Lipolase®, Lipolase( Ultra, Lipozyme®, Palatase®, Novozym® 435, Lecitase® (all available from Novo Nordisk A/S). 0 Examples of other lipases are Lumafast(, Ps. mendocian lipase from Genencor Int. Inc.; Lipomax(, Ps. pseudoalcaligenes lipase from Gist Brocades/Genencor Int. Inc.; Fusarium solani lipase (cutinase) from Unilever; Bacillus sp. lipase from Solvay enzymes. Other lipases are available from other companies. It is to be understood that also lipase variants are contemplated as the parent enzyme. 5 Examples of such are described in e.g. WO 93/01285 and WO 95/22615.
The activity of the lipase can be determined as described in "Methods of Enzymatic Analysis", Third Edition, 1984, Verlag Chemie, Weinhein, vol. 4, or as described in AF 95/5 GB (available on request from Novo Nordisk A/S).
o Parent Oxidoreductases
Parent oxidoreductases (i.e. enzymes classified under the Enzyme Classification number E.C. 1 (Oxidoreductases) in accordance with the Recommendations (1992) of the International
Union of Biochemistry and Molecular Biology (IUBMB)) include oxidoreductases within this group.
Examples include oxidoreductases selected from those classified under the Enzyme Classification (E.C.) numbers: Glycerol-3-phosphate dehydrogenase _NAD+_ (1.1.1.8), Glycerol-3 -phosphate dehydrogenase _NAD(P)+_ (1.1.1.94), Glycerol-3 -phosphate 1 -dehydrogenase _NADP_ (1.1.1.94), Glucose oxidase (1.1.3.4), Hexose oxidase (1.1.3.5), Catechol oxidase (1.1.3.14), Bilirubin oxidase (1.3.3.5), Alanine dehydrogenase (1.4.1.1), Glutamate dehydrogenase (1.4.1.2), Glutamate dehydrogenase _NAD(P)+_ (1.4.1.3), Glutamate dehydrogenase _NADP+_ (1.4.1.4), L-Amino acid dehydrogenase (1.4.1.5), Serine dehydrogenase (1.4.1.7), Valine dehydrogenase _NADP+_ (1.4.1.8), Leucine dehydrogenase (1.4.1.9), Glycine dehydrogenase (1.4.1.10), L-Amino-acid oxidase (1.4.3.2.), D-Amino-acid oxidase(1.4.3.3), L-Glutamate oxidase (1.4.3.11), Protein-lysine 6-oxidase (1.4.3.13), L-lysine oxidase (1.4.3.14), L-Aspartate oxidase (1.4.3.16), D-amino-acid dehydrogenase (1.4.99.1), Protein disulfide reductase (1.6.4.4), Thioredoxin reductase (1.6.4.5), Protein disulfide reductase (glutathione) (1.8.4.2), Laccase (1.10.3.2), Catalase (1.11.1.6), Peroxidase (1.11.1.7), Lipoxygenase (1.13.11.12), Superoxide dismutase (1.15.1.1)
Said Glucose oxidases may be derived from Aspergillus niger. Said Laccases may be derived from Polyporus pinsitus, Myceliophtora thermophila, Coprinus cinereus, Rhizoctonia solani,
Rhizoctonia praticola, Scytalidium thermophilum and Rhus vemicifera. Bilimbin oxidases may be derived from Myrothechecium verrucaria. The Peroxidase may be derived from e.g.
Soy bean, Horseradish or Coprinus cinereus. The Protein Disulfide reductase may be any of the mentioned in DK patent applications No. 768/93, 265/94 and 264/94 (Novo Nordisk A S), which are hereby incorporated as references, including Protein Disulfide reductases of bovine origin, Protein Disulfide reductases derived from Aspergillus oryzae or Aspergillus niger, and
DsbA or DsbC derived from Escherichia coli.
Specific examples of readily available commercial oxidoreductases include Gluzyme (enzyme available from Novo Nordisk A/S). However, other oxidoreductases are available from others. It is to be understood that also variants of oxidoreductases are contemplated as the parent enzyme.
The activity of oxidoreductases can be determined as described in "Methods of Enzymatic Analysis", third edition, 1984, Verlag Chemie, Weinheim, vol. 3.
Parent Carbohydrases 5 Parent carbohydrases may be defined as all enzymes capable of breaking down carbohydrate chains (e.g. starches) of especially five and six member ring structures (i.e. enzymes classified under the Enzyme Classification number E.C. 3.2 (glycosidases) in accordance with the Recommendations (1992) of the International Union of Biochemistry and Molecular Biology (IUBMB)). o Examples include carbohydrases selected from those classified under the Enzyme Classification (E.C.) numbers:
(-amylase (3.2.1.1) (-amylase (3.2.1.2), glucan l,4-(-glucosidase (3.2.1.3), cellulase (3.2.1.4), endo-l,3(4)-(-glucanase (3.2.1.6), endo-l,4-(-xylanase (3.2.1.8), dextranase (3.2.1.11), chitinase (3.2.1.14), polygalacturonase (3.2.1.15), lysozyme (3.2.1.17), (-glucosidase s (3.2.1.21), (-galactosidase (3.2.1.22), (-galactosidase (3.2.1.23), amylo-l,6-glucosidase (3.2.1.33), xylan 1 ,4-(-xylosidase (3.2.1.37), glucan endo-l,3-(-D-glucosidase (3.2.1.39), (- dextrin endo-l,6-glucosidase (3.2.1.41), sucrose (-glucosidase (3.2.1.48), glucan endo-l,3-(- glucosidase (3.2.1.59), glucan 1 ,4-(-glucosidase (3.2.1.74), glucan endo-l,6-(- glucosidase (3.2.1.75), arabinan endo-l,5-(-arabinosidase (3.2.1.99), lactase (3.2.1.108), and chitonanase 0 (3.2.1.132).
Examples of relevant carbohydrases include (-1,3-glucanases derived from Trichoderma harzianum; (-1,6-glucanases derived from a strain of Paecilomyces; (-glucanases derived from Bacillus subtilis; (-glucanases derived from Humicola insolens; (-glucan-ases derived from Aspergillus niger; (-glucanases derived from a strain of Trichoderma; (-glucanases derived 5 from a strain of Oerskovia xanthineolytica; exo-l,4-(-D-glucosidases (glucoamylases) derived from Aspergillus niger; (-amylases derived from Bacillus subtilis; (-amylases derived from Bacillus amyloliquefaciens; (-amylases derived from Bacillus stearothermophilus; (-amylases derived from Aspergillus oryzae; (-amylases derived from non-pathogenic microorganisms; (- galactosidases derived from Aspergillus niger; Pentosanases, xylanases, cellobiases, o cellulases, hemi-cellulases deliver from Humicola insolens; cellulases derived from Trichoderma reesei; cellulases derived from non-pathogenic mold; pectinases, cellulases, arabinases, hemi-celluloses derived from Aspergillus niger; dextranases derived from
Penicillium lilacinum; endo-glucanase derived from non-pathogenic mold; pullulanases derived from Bacillus acidopullyticus; (-galactosidases derived from Kluyveromyces fragilis; xylanases derived from Trichoderma reesei;
Specific examples of readily available commercial carbohydrases include Alpha-Gal(, Bio- 5 Feed( Alpha, Bio-Feed( Beta, Bio-Feed( Plus, Bio-Feed( Plus, Novozyme® 188, Carezyme®, Celluclast®, Cellusoft®, Ceremyl®, Citrozym(, Denimax(, Dezyme(, Dextrozyme(, Finizym®, Fungamyl(, Gamanase(, Glucanex®, Lactozym®, Maltogenase(, Pentopan(, Pectinex(, Promozyme®, Pulpzyme(, Novamyl(, Termamyl®, AMG (Amyloglucosidase Novo), Maltogenase®, Aquazym®, Natalase( (all enzymes available from Novo Nordisk o A S). Other carbohydrases are available from other companies.
It is to be understood that also carbohydrase variants are contemplated as the parent enzyme. The activity of carbohydrases can be determined as described in "Methods of Enzymatic Analysis", third edition, 1984, Verlag Chemie, Weinheim, vol. 4.
s Parent Transferases
Parent transferases (i.e. enzymes classified under the Enzyme Classification number E.C. 2 in accordance with the Recommendations (1992) of the International Union of Biochemistry and Molecular Biology (IUBMB)) include transferases within this group. The parent transferases may be any transferase in the subgroups of transferases: transferases 0 transferring one-carbon groups (E.C. 2.1); transferases transferring aldehyde or residues (E.C 2.2); acyltransferases (E.C. 2.3); glucosyltransferases (E.C. 2.4); transferases transferring alkyl or aryl groups, other that methyl groups (E.C. 2.5); transferases transferring nitrogeneous groups (2.6). In a prefened embodiment the parent transferase is a transglutaminase E.C 2.3.2.13 (Protein- 5 glutamine (-glutamyltransferase) .
Transglutaminases are enzymes capable of catalyzing an acyl transfer reaction in which a gamma-carboxyamide group of a peptide-bound glutamine residue is the acyl donor. Primary amino groups in a variety of compounds may function as acyl acceptors with the subsequent formation of monosubstituted gamma-amides of peptide-bound glutamic acid. When the o epsilon-amino group of a lysine residue in a peptide-chain serves as the acyl acceptor, the transferases form intramolecular or intermolecular gamma-glutamyl-epsilon-lysyl crosslinks.
Examples of transglutaminases are described in the pending DK patent application no. 990/94 (Novo Nordisk A/S).
The parent transglutaminase may be of human, animal (e.g. bovine) or microbial origin. Examples of such parent transglutaminases are animal derived Transglutaminase, FXUIa; microbial transglutaminases derived from Physarum polycephalum (Klein et al., Journal of Bacteriology, Vol. 174, p. 2599-2605); transglutaminases derived from Streptomyces sp., including Streptomyces lavendulae, Streptomyces lydicus (former Streptomyces libani) and Streptoverticillium sp., including Streptoverticillium mobaraense, Streptoverticillium cin- namoneum, and Streptoverticillium griseocarneum (Motoki et al., US 5,156,956; Andou et al., US 5,252,469; Kaempfer et al., Journal of General Microbiology, Vol. 137, p. 1831-1892; Ochi et al, International Journal of Sytematic Bacteriology, Vol. 44, p. 285-292; Andou et al., US 5,252,469; Williams et al., Journal of General Microbiology, Vol. 129, p. 1743-1813). It is to be understood that also transferase variants are contemplated as the parent enzyme. The activity of transglutaminases can be determined as described in "Methods of Enzymatic Analysis", third edition, 1984, Verlag Chemie, Weinheim, vol. 1-10.
Parent Phytases
Parent phytases are included in the group of enzymes classified under the Enzyme Classification number E.C. 3.1.3 (Phosphoric Monoester Hydrolases) in accordance with the Recommendations (1992) of the International Union of Biochemistry and Molecular Biology
(IUBMB)).
Phytases are enzymes produced by microorganisms, which catalyse the conversion of phytate to inositol and inorganic phosphorus.
Phytase producing microorganisms comprise bacteria such as Bacillus subtilis, Bacillus natto and Pseudomonas; yeasts such as Saccharomyces cerevisiae; and fungi such as Aspergillus niger, Aspergillus ficuum, Aspergillus awamori, Aspergillus oryzae, Aspergillus teneus or
Aspergillus nidulans, and various other Aspergillus species).
Examples of parent phytases include phytases selected from those classified under the Enzyme
Classification (E.C.) numbers: 3-phytase (3.1.3.8) and 6-phytase (3.1.3.26). The activity of phytases can be determined as described in "Methods of Enzymatic Analysis", third edition, 1984, Verlag Chemie, Weinheim, vol. 1-10, or may be measured according to the method described in EP-A1-0 420 358, Example 2 A.
Lyases
Suitable lyases include Polysaccharide lyases: Pectate lyases (4.2.2.2) and pectin lyases
(4.2.2.10), such as those from Bacillus licheniformis disclosed in WO 99/27083.
Isomerases:
Protein Disulfide Isomerase.
Without being limited thereto suitable protein disulfide isomerases include PDIs described in WO 95/01425 (Novo Nordisk A/S) and suitable glucose isomerases include those described in Biotechnology Letter, Vol. 20, No 6, June 1998, pp. 553-56.
Contemplated isomerases include xylose/glucose Isomerase (5.3.1.5) including Sweetzyme®.
Identifying areas of interest for introduction of modifications
The methods of this invention are especially suitable when testing compounds that are being modified with respect to their allergenicity.
Such modification of a test compound to affect its immunogenecity could be by mutation of a protein allergen in its immunoglobulin-specific epitopes. The location of these epitopes can be determined by several techniques such as those disclosed by WO 92/10755 (by U. Løvborg), by Walshet et al, J. Immunol. Methods, vol. 121, 1275-280, (1989), and by Schoofs et al. J. Immunol, vol. 140, 611-616, (1987). A prefened method for identification of epitopes is by screening a random peptide library with antibodies (e.g. IgE or IgG antibodies), selecting high-binding peptides, obtaining the sequence these, and aligning the high-binding peptide sequences to identify a consensus sequence. These consensus sequences, in turn are compared with the sequence and 3D structure of a relevant protein, which is desired to mutate for reduction of immunogenicity, in order to identify the linear and structural epitopes of that protein.
In an even more prefened method, the identification of epitope(s) may be achieved by screening of phage display libraries. The principle behind phage display is that a heterologous DNA sequence can be inserted in the gene coding for a coat protein of the phage. The phage will make and display the hybrid protein on its surface where it can interact with specific target agents. Such target agent may be antigen-specific antibodies. It is therefore possible to select specific phages that display antibody-binding peptide sequences. The displayed peptides can be of predetermined lengths, for example 9 amino acids long, with randomized sequences, resulting in a random peptide display package library. Thus, by screening for antibody binding, one can isolate the peptide sequences that have the highest affinity for the particular antibody used. The peptides of the hybrid proteins of the specific phages which bind protein- specific antibodies define the epitopes of that particular protein. The conesponding residues of the parent protein can be found by aligning the selected peptide sequences resulting in epitope patterns, and compare these with the amino acid sequence and 3-dimensional structure of the parent protein.
When the epitope(s) have been identified, a protein variant exhibiting a reduced immunogenicity may be produced by changing the identified epitope pattern of the parent protein by genetic engineering of a DNA sequence encoding the parent protein. It is commonly found, that amino acids sunounding B- and T-cell epitopes can affect binding of the antibodies or T-cell receptors to the antigen. To anticipate this possibility, an epitope area was defined on the 3-dimensional structure of the protein of interest, and genetic engineering of any amino acid within the epitope area is considered to be within the scope of using the identified epitope to generate variants with low immunogenecity.
Generating a diversified library
In order to generate protein variants, more than one amino acid residue may be substituted, added or deleted, these amino acids preferably being located in different epitope areas. In that case, it may be difficult to assess a priori how well the functionality of the protein is maintained while antigenicity is reduced, especially since the possible number of combinations of mutations become very large, even for a small number of mutations. In that
case, it will be an advantage, to establish a library of diversified mutants each having one or more changed amino acids introduced and selecting those variants, which show good retention of function and at the same time a significant reduction in antigenicity.
5 A diversified library can be established by a range of techniques known to the person skilled in the art (Reetz MT; Jaeger KE, in 'Biocatalysis - from Discovery to Application' edited by Fessner WD, Vol. 200, pp. 31-57 (1999); Stemmer, Nature, vol. 370, p.389-391, 1994; Zhao and Arnold, Proc. Natl. Acad. Sci., USA, vol. 94, pp. 7997-8000, 1997; or Yano et al., Proc. Natl. Acad. Sci., USA, vol. 95, pp 5511-5515, 1998). o These include, but are not limited to, 'spiked mutagenesis', in which certain positions of the protein sequence are randomized by earring out PCR mutagenesis using one or more oligonucleotide primers which are synthesized using a mixture of nucleotides for certain positions (Lanio T, Jeltsch A, Biotechniques, Vol. 25(6), 958,962,964-965 (1998)). The mixtures of oligonucleotides used within each triplet can be designed such that the s conesponding amino acid of the mutated gene product is randomized within some predetermined distribution function. Algorithms exist, which facilitate this design (Jensen LJ et al., Nucleic Acids Research, Vol. 26(3), 697-702 (1998)).
In an embodiment substitutions are found by a method comprising the following steps: 1) a o range of substitutions, additions, and/or deletions are listed encompassing several epitope areas, 2) a library is designed which introduces a randomized subset of these changes in the amino acid sequence into the target gene, e.g. by spiked mutagenesis, 3) the library is expressed, and prefened variants are selected. In another embodiment, this method is supplemented with additional rounds of screening and/or family shuffling of hits from the first 5 round of screening (J.E. Ness, et al, Nature Biotechnology, vol. 17, pp. 893-896, 1999) and/or combination with other methods of reducing immunogenicity by genetic means (such as that disclosed in WO92/10755).
The library may be designed, such that at least one amino acid of the epitope area is o substituted. In a prefened embodiment at least one amino acid of the epitope itself is changed.
The library may be biased such that towards introducing an amino acid of different size, hydrophilicity, and/or polarity relative to the original one of the 'protein backbone'. For
example changing a small amino acid to a large amino acid, a hydrophilic amino acid to a hydrophobic amino acid, a polar amino acid to a non-polar amino acid or a basic to an acidic amino acid. Other changes may be the addition or deletion of at least one amino acid of the epitope area, preferably deleting an anchor amino acid. Furthermore, substituting some amino acids and deleting or adding others may change an epitope.
In another embodiment, the library is designed, such that recognition sites for post- translational modifications are introduced in the epitope areas, and the library is expressed in a suitable host organism capable of the conesponding post-translational modification. These post-translational modifications may serve to shield the epitope and hence lower the immunogenicity of the protein variant relative to the protein backbone. Post-translational modifications include glycosylation, phosphorylation, N-terminal processing, acylation, ribosylation and sulfatation. A good example is N-glycosylation. N-glycosylation is found at sites of the sequence Asn-Xaa-Ser, Asn-Xaa-Thr, or Asn-Xaa-Cys, in which neither the Xaa residue nor the amino acid following the tri -peptide consensus sequence is a pro line (T. E. Creighton, 'Proteins - Structures and Molecular Properties, 2nd edition, W.H. Freeman and Co., New York, 1993, pp. 91-93). It is thus desirable to introduce such recognition sites in the sequence of the backbone protein. The specific nature of the glycosyl chain of the glycosylated protein variant may be linear or branched depending on the protein and the host cells. Another example is phosphorylation: The protein sequence can be modified so as to introduce serine phophorylation sites with the recognition sequence arg-arg-(xaa)n-ser (where n = 0, 1, or 2), which can be phosphorylated by the cAMP-dependent kinase or tyrosine phosphorylation sites with the recognition sequence -lys/arg - (xaa)3 - asp/glu- (xaa)3 - tyr, which can usually be phophorylated by tyrosine-specific kinases (T.E. Creighton, "Proteins- Stmctures and molecular properties", 2nd ed., Freeman, NY, 1993).
Covalent conjugation to amino acids in the epitope area.
In yet another embodiment, one can design the libarary, such that amino acids suitable for chemical modification are substituted for existing ones in the epitope areas. The protein variant can then be conjugated to activated polymers Which amino acids to substitute and/or insert depends in principle on the coupling chemistry to be applied. The chemistry for
preparation of covalent bioconjugates can be found in "Bioconjugate Techniques", Hermanson, G.T. (1996), Academic Press Inc., which is hereby incorporated as reference (see below). It is prefened to make conservative substitutions in the polypeptide when the polypeptide has to be conjugated, as conservative substitutions secure that the impact of the substitution on the polypeptide structure is limited.In the case of providing additional amino groups this may be done by substitution of arginine to lysine, both residues being positively charged, but only the lysine having a free amino group suitable as an attachment groups.In the case of providing additional carboxylic acid groups the conservative substitution may for instance be an asparagine to aspartic acid or glutamine to glutamic acid substitution. These residues resemble each other in size and shape, except from the carboxylic groups being present on the acidic residues. In the case of providing SH-groups the conservative substitution may be done by changing threonine or serine to cysteine.
Diversity in the protein variant library can be generated at the DNA triplet level, such that individual codons are variegated e.g. by using primers of partially randomized sequence for a PCR reaction. Further, several techniques have been described, by which one can create a library with such diversity at several locations in the gene, which are too far apart to be covered by a single (spiked) oligonucleotide primer. These techniques include the use of in vivo recombination of the individually diversified gene segments as described in WO 97/07205 on page 3, line 8 to 29 or by using DNA shuffling techniques to create a library of full length genes that combine several gene segments each of which are diversified e.g. by spiked mutagenesis (Stemmer, Nature 370, pp. 389-391, 1994 and US 5,605,793 and 5,830,721). In the latter case, one can use the gene encoding the "protein backbone" as a template double-stranded polynucleotide and combining this with one or more single or double-stranded oligonucleotides as described in claim 1 of US 5,830,721. The single- stranded oligonucleotides could be partially randomized during synthesis. The double- stranded oligonucleotides could be PCR products incorporating diversity in a specific region. In both cases, one can dilute the diversity with conesponding segments containing the sequence of the backbone protein in order to limit the number of changes that are on average introduced. As mentioned above, methods have been established for designing the ratios of nucleotides (A; C; T; G) used at a particular codon during primer synthesis, so as to approximate a desired frequency distribution among a set of desired amino acids at that
particular codon. This allows one to bias the partially randomized mutagenesis towards e.g. introduction of post-translational modification sites, chemical modification sites, or simply amino acids that are different from those that define the epitope or the epitope area. One could also approximate a sequence in a given location or epitope area to the conesponding location on a homologous, human protein.
When one uses protein engineering to eliminate epitopes, it is indeed possible that new epitopes are created, or existing epitopes are duplicated. To reduce this risk, one can map the planned mutations at a given position on the 3-dimensional structure of the protein of interest, and control the emerging amino acid constellation against a database of known epitope patterns, to rule out those possible replacement amino acids, which are predicted to result in creation or duplication of epitopes. Thus, risk mutations can be identified and eliminated by this procedure, thereby reducing the number of mutations at each position, and hence reducing the library size. Occasionally, one would be interested in testing a library that combines a number of known mutations in different locations in the primary sequence of the 'protein backbone'. These could be introduced post-translational or chemical modification sites, or they could be mutations, which by themselves had proven beneficial for one reason or another (e.g. decreasing antigenicity, or improving specific activity, performance, stability, or other characteristics). In such cases, it may be desirable to create a library of diverse combinations of known sequences. For example if 12 individual mutations are known, one could combine (at least) 12 segments of the 'protein backbone' gene in which each segment is present in two forms: one with and one without the desired mutation. By varying the relative amounts of those segments, one could design a library (of size 212) for which the average number of mutations per gene can be predicted. This can be a useful way of combining elements that by themselves give some, but not sufficient effect, without resorting to very large libraries, as is often the case when using 'spiked mutagenesis'..
Host cells, culturing, and sampling
As described above, any number of host cells can be used to perform the method of the invention. Preferably the host cells are of microbial origin, preferably bacterial, yeast, or fungal. Even more preferably the host cells are chosen from the group consisting of Escherichia coli, Bacillus subtilis, Bacillus licheniformis, Bacillus clausii, Aspergillus niger, Aspergillus oryzae, Aspergillus nidulans, and Saccharomyces cerevisiae.
The diversified library is prepared as a DNA library of genes encoding variants of the relevant protein backbone. This DNA library is then transformed into the host cells by using any of the techniques known in the art and described above. Typically, one would want to culture the cells in different positions of a spatial anay, necessitating the distribution of individual clones into each position of the anay. This is ideally done in such a way that each position is occupied by exactly one cell. In practice, however, the number of cells at each position will follow a probability distribution. Hence, in a prefened embodiment, the average number of cells per well is between 0,2 and 1 cell. In a more prefened embodiment, the number of cells per well is optimized such that the highest density of anay positions occupied by exactly one cell is obtained. The number of cells at each position can be controlled by dilution. The dilution that most closely approximates one cell per position is often termed 'the limiting dilution'.
Steps (iii) through (vi) of the present invention can be performed in many ways. Typically, the culturing and the sampling of host cells takes place using a spatial anay. The spatial anay can take on any physical form whatsoever, that enables the culturing or assaying of several samples at once, without one sample contaminating another. Examples of prefened spatial anays are different kinds of microtiter-plates with any number of wells, such as 96 or 384, and of any kind of material, as well as positions in a High Performance Liquid Chromatography (HPLC) autosampler device. Any kind of physical anangement which allows the unambiguous identification of the samples by a number or a position in the anay. Even samples placed as drops on a surface in a specific recorded pattern, the surface being of a solid material or of more complex nature such as a textile or a tissue, e.g. cotton, wool, paper, or cellulose.
A prefened embodiment relates to a method, wherein the spatial anay of is a microtiter plate, a solid surface or a textile surface.
A way of carrying out step (iv) of the first aspect of the invention could be to take a sample 5 from each position of the spatial anay, e.g. from a supernatant or cell culture, and transfer this to another spatial anay for further testing or assaying for production of the molecule of interest. The second spatial anay may be identical to the first one used in the specific method, but may also be of any other kind that fulfills the above mentioned criteria, such as a microtiter plate, a solid surface, a textile, any material etc. o Accordingly a prefened embodiment relates to a method, wherein after step (iv) a sample is transfened from each position in the spatial anay to a position in a second spatial anay which is then used onwards in the method, preferably the second anay is a microtiter plate. In one aspect of the cunent invention, the individual cells/clones are grown within microenvironments in each position of the spatial anay. Such microenvironments are initally s sterile beads or balls of any material that allows growth of the clone, preferably beads comprising agarose, alginate, polysaccharide, carbohydrate, alginate, canageenan, chitosan, cellulose, pectin, dextran or polyacrylamide, all allowing diffusion to/from each micro- environment.
0 Accordingly a prefened embodiment relates to a method of the first aspect, wherein each position in the spatial anay is occupied by a bead comprising one cell, preferably the bead is an agarose-bead.
The assaying of the invention for production of the molecule of interest can be done in a great 5 number of ways. As indicated, some kinds of spatial anays like microtiter plates, can sometimes be assayed directly, or samples can be taken from each position and tranfened to another spatial anay for assaying according to step (v) and (vi) or according to any of a number of techniques well known in the art, such as enzyme activity, receptor binding, and many others; common to these assays is that a measurement is taken of a detectable property o e.g. fluorescence, luminescence, absorption. The method of the invention can be performed with any number of these assays, and consequently the molecules assayed for in step (v) and (vi) can be obtained in a number of ways, depending on the host cell construct. The molecule
of interest may be secreted by the host cell into the supernatant, or the molecule may remain intracellular, in which case lysis of the host cells may release the molecule. A prefened embodiment relates to a method, wherein the molecule of interest in (v) and (vi) is assayed in either whole broth, supernatant of cells that secrete the molecule, a lysate of cells that produce the molecule, or is assayed while still inside cells that produce the molecule.
In another embodiment, the sample is purified by a size separation process, such as the membrane processes filtration and dialysis, prior to the analyses of step (v) and (vi). This will serve to reduce interference from particulate matter or high molecular mass compounds from the cells (which are larger than the soluble protein variants) or from the small molecule metabolites, substrate compounds, or degradation products (which are smaller than the protein variants). This size separation can be achieved, e.g. by using microtiter well inserts (e.g. Nunc TC), vacuum filtration microtiter plates (e.g. Qiagen, QIAwell) or pumping or centrifugation devices.
Extra steps
In one aspect of the present invention the host cells may be first selected based on a functional screen, and then re-cultured, sampled, and analysed for antibody binding capacity and functionality. The initial functional selection can be done using an agar plate assay to analyse colonies from the transformed cells and select for those resulting in halo formation (see for example Ness, J.E:, et al., Nature Biotechn., 17, pp 893-896, 1999). Then, host cells expressing functional protein variants may be selected, preferably by automated colony picking, before the antibody binding capacity of the protein variant is assayed. This mode can increase the throughput and quality of the screening setup in cases where the antibody binding capacity assay is slower, more cumbersome, more expensive, or less accurate than the functionality assay.
In another embodiment of the invention, protein variant is exposed to adverse conditions prior to analyzing the functionality, in order to gauge also the stability of the protein variant.
Alternatively, the functionality is determined twice with the protein variant being exposed to adverse conditions for a period of time in between the two determinations of functionality. From these measurements the stability of the protein variant at the particular set of adverse conditions can be determined. These adverse conditions could be characterized by increased temperature, increased or decreased pH, the presence of certain metal ions or of metal chelators such as EDTA. They could also be the presence of surfactant molecules, such as those used in detergents, skin cream formulations, hand dish washing compositions or other applications; the presence of proteases or other degradative enzymes; or they could be the presence of enzyme inhibitors. In this embodiment, one records the stability as an additional parameter, which may help in selecting the protein variants for further culturing and analysis.
In another embodiment, the protein variant of the sample (step iv) is immobilized to a solid material. This will be an advantage, for instance by facilitating removal of impurities, chemical modification of the protein variant, and assessment of the antibody binding capacity of the protein variant. The solid material could be the well surface of a microtiterplate, suspended beads, the pin of dipstick devices or others. Immobilization can be either nonspecific by hydrophobic interactions, ionic interactions, or chemical coupling, or it can be more specific, such as non-covalent coupling to immobilized binding partners such as antibodies, enzyme inhibitors, or substrate mimics. In a prefened embodiment, the protein backbone has been modified genetically with an N-terminal or C-terminal extension comprising a high affinity tag for a compound, which can be immobilized. An example is a polyhistidine tag, which binds with high affinity to Nickel, which in turn has been bound to an acid such as nitrilotriacetic acid which has been immobilized to the solid phase. Other examples of an affinity tags are the cellulose binding domains of bacteria and fungi, which bind with high affinity to cellulose (such as Avicell), also when fused onto heterogeneous proteins, calmodulin binding domains, S-tag or FLAG peptides that bind to specific antibodies etc.
In a prefened embodiment, the binding to the solid phase is reversible, such that the protein variant can be eluted into solution when exposed to certain conditions such as e.g. high salt concentrations, high pH, or high competitor concentration.
In another embodiment, efficient HTS is achieved by devicing a simple yet accurate determination of the amount of a specific active molecule produced by the individual clone. One solution is that when the concentration of the specific molecule in a position of the spatial anay has been determined, this information can be used to determine the specific activity from the total activity determined in that position; alternatively, the information can be used to adjust the input of the molecule into the activity assay. A second solution is to use immobilization of the protein variant as a means to dose the subsequent assays with a known and/or constant amount of protein variant. This requires that the sample contain more protein variant than the binding capacity of the immobilization step. Alternatively, the assay can be configured in such a way that it is insensitive to the concentration of the molecule.
Chemical conjugation
In one embodiment, the protein variants are being modified by chemical conjugation prior to the analyses of step (v) and (vi). For this, the protein variant needs to be incubate with an active or activated polymer and subsequently separated from the unreacted polymer. This can conveniently be done using the immobilized protein variants, which can easily be exposed to different reaction environments and washes.
In the case were polymeric molecules are to be conjugated with the polypeptide in question and the polymeric molecules are not active they must be activated by the use of a suitable technique. It is also contemplated according to the invention to couple the polymeric molecules to the polypeptide through a linker. Suitable linkers are well-known to the skilled person.
Methods and chemistry for activation of polymeric molecules as well as for conjugation of polypeptides are intensively described in the literature. Commonly used methods for activation of insoluble polymers include activation of functional groups with cyanogen bromide, periodate, glutaraldehyde, biepoxides, epichlorohydrin, divinylsulfone, carbodiimide, sulfonyl halides, trichlorotriazine etc. (see R.F. Taylor, (1991), "Protein immobilisation. Fundamental and applications", Marcel Dekker, N.Y.; S.S. Wong, (1992), "Chemistry of Protein Conjugation and Crosslinking", CRC Press, Boca Raton; G.T.
Hermanson et al., (1993), "Immobilized Affinity Ligand Techniques", Academic Press, N.Y.).
Some of the methods concern activation of insoluble polymers but are also applicable to activation of soluble polymers e.g. periodate, trichlorotriazine, sulfonylhalides, divinylsulfone, carbodiimide etc. The functional groups being amino, hydroxyl, thiol, carboxyl, aldehyde or sulfydryl on the polymer and the chosen attachment group on the protein must be considered in choosing the activation and conjugation chemistry which normally consist of i) activation of polymer, ii) conjugation, and iii) blocking of residual active groups.
In the following a number of suitable polymer activation methods will be described shortly.
However, it is to be understood that also other methods may be used. Coupling polymeric molecules to the free acid groups of polypeptides may be performed with the aid of diimide and for example amino-PEG or hydrazino-PEG (Pollak et al., (1976), J.
Am. Chem. Soc, 98, 289-291) or diazoacetate/amide (Wong et al., (1992), "Chemistry of
Protein Conjugation and Crosslinking", CRC Press).
Coupling polymeric molecules to hydroxy groups is generally very difficult as it must be performed in water. Usually hydrolysis predominates over reaction with hydroxyl groups.
Coupling polymeric molecules to free sulfhydryl groups can be achieved with special groups like maleimido or the ortho-pyridyl disulfide. Also vinylsulfone (US patent no. 5,414,135,
(1995), Snow et al.) has a preference for sulfhydryl groups but is not as selective as the other mentioned. Accessible Arginine residues in the polypeptide chain may be targeted by groups comprising two vicinal carbonyl groups.
Techniques involving coupling of electrophilically activated PEGs to the amino groups of Lysines may also be useful. Many of the usual leaving groups for alcohols give rise to an amine linkage. For instance, alkyl sulfonates, such as tresylates (Nilsson et al., (1984), Methods in Enzymology vol. 104, Jacoby, W. B., Ed., Academic Press: Orlando, p. 56-66; Nilsson et al., (1987), Methods in Enzymology vol. 135; Mosbach, K., Ed.; Academic Press: Orlando, pp. 65-79; Scouten et al., (1987), Methods in Enzymology vol. 135, Mosbach, K., Ed., Academic Press: Orlando, 1987; pp 79-84; Crossland et al., (1971), J. Amr. Chem. Soc. 1971, 93, pp. 4217-4219), mesylates (Harris, (1985), supra; Harris et al., (1984), J. Polym. Sci. Polym. Chem. Ed. 22, pp 341-352), aryl sulfonates like tosylates, and para-nitrobenzene sulfonates can be used.
Organic sulfonyl chlorides, e.g. Tresyl chloride, effectively converts hydroxy groups in a number of polymers, e.g. PEG, into good leaving groups (sulfonates) that, when reacted with nucleophiles like amino groups in polypeptides allow stable linkages to be formed between polymer and polypeptide. In addition to high conjugation yields, the reaction conditions are in 5 general mild (neutral or slightly alkaline pH, to avoid denaturation and little or no disruption of activity), and satisfy the non-destructive requirements to the polypeptide. Tosylate is more reactive than the mesylate but also less stable decomposing into PEG, dioxane, and sulfonic acid (Zalipsky, (1995), Bioconjugate Chem., 6, 150-165). Epoxides may also been used for creating amine bonds but are much less reactive than the abovementioned lo groups.
Converting PEG into a chloroformate with phosgene gives rise to carbamate linkages to Lysines. Essentially the same reaction can be carried out in many variants substituting the chlorine with N-hydroxy succinimide (US patent no. 5,122,614, (1992); Zalipsky et al., (1992), Biotechnol. Appl. Biochem., 15, p. 100-114; Monfardini et al., (1995), Bioconjugate is Chem., 6, 62-69, with imidazole (Allen et al., (1991), Carbohydr. Res., 213, pp 309-319), with para-nitrophenol, DMAP (EP 632 082 Al, (1993), Looze, Y.) etc. The derivatives are usually made by reacting the chloroformate with the desired leaving group. All these groups give rise to carbamate linkages to the peptide. Furthermore, isocyanates and isothiocyanates may be employed, yielding ureas and thioureas,
2 o respectively.
Amides may be obtained from PEG acids using the same leaving groups as mentioned above and cyclic imid thrones (US patent no. 5,349,001, (1994), Greenwald et al.). The reactivity of these compounds are very high but may make the hydrolysis to fast. PEG succinate made from reaction with succinic anhydride can also be used. The hereby
25 comprised ester group make the conjugate much more susceptible to hydrolysis (US patent no. 5,122,614, (1992), Zalipsky). This group may be activated with N-hydroxy succinimide. Furthermore, a special linker can be introduced. The most well studied being cyanuric chloride (Abuchowski et al, (1977), J. Biol. Chem., 252, 3578-3581; US patent no. 4,179,337, (1979), Davis et al.; Shafer et al., (1986), J. Polym. Sci. Polym. Chem. Ed., 24,
30 375-378.
Coupling of PEG to an aromatic amine followed by diazotation yields a very reactive diazonium salt, which can be reacted with a peptide in situ. An amide linkage may also be
obtained by reacting an azlactone derivative of PEG (US patent no. 5,321,095, (1994), Greenwald, R. B.) thus introducing an additional amide linkage.
As some peptides do not comprise many Lysines it may be advantageous to attach more than one PEG to the same Lysine. This can be done e.g. by the use of l,3-diamino-2-propanol. PEGs may also be attached to the amino-groups of the enzyme with carbamate linkages (WO 95/11924, Greenwald et al.). Lysine residues may also be used as the backbone. The coupling technique used in the examples is the N-succinimidyl carbonate conjugation technique descried in WO 90/13590 (Enzon).
In a prefened embodiment, the activated polymer is methyl-PEG which has been activated by N-succinimidyl carbonate as described WO 90/13590. The coupling can be carried out at alkaline conditions in high yields.
For coupling of polymers to the protein variants, it is prefened to use conditions similar to those described in WO96/17929 and WO99/00489 (Novo Nordisk A S) e.g. mono or bis activated PEG's of molecular weight ranging from 100 to 5000 Da. For instance, a methyl- PEG 350 could be activated with N-succinimidyl carbonate and incubated with protein variant at a molar ratio of more than 5 calculated as equivalents of activated PEG divided by moles of lysines in the protein of interest. For coupling to immobilized protein variant, the PEG:protein ratio should be optimized such that the PEG concentration is low enough for the buffer capacity to maintain alkaline pH throughout the reaction; while the PEG concentration is still high enough to ensure sufficient degree of modification of the protein. Further, it is important that the activated PEG is kept at conditions that prevent hydrolysis (i.e. dissolved in acid or solvents) and diluted directly into the alkaline reaction buffer. It is essential that primary amines are not present other than those occurring in the lysine residues of the protein. This can be secured by washing thoroughly in borate buffer. The reaction is stopped by separating the fluid phase containing unreacted PEG from the solid phase containing protein and derivatized protein. Optionally, the solid phase can then be washed with tris buffer, to block any unreacted sites on PEG chains that might still be present.
Determining antibody binding capacity:
In step (v) of the first aspect of the invention, a sample is analysed to determine the antibody binding capacity of its variant protein. The antibodies used for this analysis can be in the form of serum isolated from an animal such as rabbit, mouse, rat, guinea pig, sheep etc., which has previously been exposed to the protein backbone. Optionally, the serum can be human serum from a volunteer who is known to have a history of exposure to the protein backbone. Preferably, serum samples are pooled from several animal or human donors to achieve an individual-independent result of the screening. Alternatively, the serum can be purified (e.g. by caprylic acid precipitation and DEAE chromatography) to achieve an immunoglobulin G fraction, or the serum can be purified using a Protein A column and/or an affinity column with anti-IgE (Fcε) antibodies, to prepare an IgE fraction. These and other methods of antibody purification are described in Harlow and Lane, "Antibodies - A laboratory manual", Cold Spring Harbor Laboratory, 1988 and in Catty and Raykundalia: Production and Quality Control of Polyclonal Antibodies in "Antibodies - a practical approach" vol. 1, IRL Press, Oxford, 1988.
When the objective of using the HTS method of this invention is to reduce allergenicity of the protein variants, the IgE-purification method is prefened, and even more preferable is an embodiment in which the experimental animal has been intratracheally exposed to the protein backbone and has been shown to develop antigen-specific IgE antibodies, or if the human volunteers have a history of allergenic sensitisation to the protein backbone.
Whether an IgG, IgE, or other immunoglobulin fraction is used, it is preferable to increase the specificity of the antibody binding analysis by further purification. This can be in the form of affinity purification using a solid phase of immobilized antigen (e.g. the protein backbone). Immobilization can be by chemical conjugation, specific binding to a ligand or an antibody, or by binding through a fusion tag such as a polyhistidine tage, a FLAG peptide or the like. Antibodies are bound and eluted (e.g. using 1 M propionic acid) resulting in a "specific polyclonal antibody preparation" (see Arvieuz and Williams: Immunoaffinity
Chromatography in "Antibodies - a practical approach" vol. 1, IRL Press, Oxford, 1988). In
the case of using a protease antigen it may be desirable to reduce or eliminate protease activity during the antibody purification step. This can be done by several methods, including binding to an immobilized inhibitor (such as bacitracin in the case of subtilisins), using a chemically inactivated protease backbone (e.g. by PMSF treatment), or by using a mutated protease (e.g. 5 by converting the catalytically active serine to an alanine). In the latter case, the antibodies are raised against the mutated version of the protein backbone, while the library is diversified using the active protein backbone as a template, and this embodiment of the method is also considered an aspect of the present invention.
o In another embodiment, the antibodies can be purified using an epitope-specific ligand. In the case of linear epitopes, this can be in the form of a peptide fragment of the protein backbone, while in the case of a structural epitope this could be in the form of a peptide (or peptide phage membrane protein fusion) which has been isolated by a peptide display library using antigen-specific polyclonal antibodies, as described above. Such purification schemes lead to s "monospecific antibodies" which are useful for the cunent invention.
In another embodiment, the antibodies can be monoclonal antibodies, each of which are considered a subgroup of "monospecific antibodies" as they have only one binding speicificity. The hybridoma clones can be selected based on specificity for the entire protein o backbone or for specificity for an epitope (as described above). In either case, the epitope specificity can be assessed for instance by standard immunoassays using the isolated antibody- binding peptides in order to assign the specificity of each hybridoma clone to a particular epitope pattern. This way, one can create libraries with variation in a single epitope and assay it with one or more monoclonal antibodies specific for that particular epitope in order to get a 5 very specific response.
Further, the antibodies, whether specific polyclonal, monospecific, or monoclonal, can be labelled to allow detection. Also, secondary antibodies directed against the primary antibodies can be labelled. The label can be a chemically bound compound such as peroxidase, o streptavidin, alkaline phosphatase, fluorescent or luminescent compounds or others, or it can be a fusion tag such as a polyhistidine tag, a FLAG peptide, an S-tag or other fusion peptides for which there are or can be made specific antibodies.
When assessing antibody binding capacity of a protein variant sampled directly from a cell culture supernatant, there will be many components of the supernatant, which may interfere with the antibody-binding assay. This background interference can be reduced by thorough purification of the protein backbone prior to sensitisation of the test animal, or by several other methods. One such method is to purify the antibodies on a column with immobilized 'cellular impurities' obtained by culturing a strain of host cells which have not been transformed or which have been transformed with a vector that does not contain any protein backbone or protein variant, as described (Naver and Løvborg, Scand. J. Immunol., 41, pp. 443-448, 1995). Another method to reduce background interference is to raise the antibodies against a protein backbone, which has been expressed in an organism or strain different from the host cells used for expression of the diversified libraries. One could use E.coli instead of bacillus, or even different strains of Bacillus (e.g. B. subtilis vs. B .licheniformis) may be sufficiently different to ensure that polyclonal antibodies are specific for impurities from one strain, but not from the other. A third method is to use immobilization of the protein variants (as described above) to facilitate removal of cell culture supernatant impurities.
In one aspect of the invention, the antibody binding assay is multivalent in nature, i.e. it depends on multivalent or bivalent interactions between antibody and antigen. Examples of such assay formats are agglutination assays and assays based on 'passive immunization' of effector cells, e.g. mast cells or basophiles, with IgE antibodies and detection of cell-specific responses to antigen-induced aggregation of the cell-surface bound IgE molecules (Skov, PS et al, Pediatr. Allergy Immunol., 8, pp. -156-158, 1995; Diamant and Pratkar, Int. Archs. Allergy appl. Immun. 67, pp. 13-17, 1982). This aspect can be advantageous when several epitopes are diversified in the same library.
In the first aspect of the invention, the antibody binding capacity is determined in step (v). In order to achieve high throughput of the screening method, it is desirable to use an antibody binding analysis that requires few dosages, preferably only one dosage of protein variant, and which in other ways is designed to give the highest likelihood of identifying low immunogenic protein variants. These variations in design of the antibody binding analysis are constitutes several embodiments of the invention, as described in the following.
Immobilized form
In one embodiment, in which the protein variants have been immobilized to a solid phase for 5 analysis, the binding can be determined using a labelled primary antigen-specific antibody or alternatively by using a primary antigen-specific antibody and a labelled secondary species- specific anti-Ig antibody. Immobilization of the protein variant makes it easier to change the reaction medium several times, introduce washing steps etc. In one aspect of this embodiment, the antibody-binding can be determined using competitive antigen (e.g. protein backbone) and o labelled antibodies in the solution.
Soluble form
The protein variants may be analysed for antibody binding capacity directly, or they may have been immobilized to a solid phase and then eluted (as described above). In either case, the s antibody binding capacity is analysed with the protein variant in solution.
In one embodiment, the antibodies have been coated onto a solid phase (such as the surface of wells in a microtiter plate. Thus, protein variants bind to (or be captured by) the immobilized antibodies at the surface, and the supernatant contain only antigens that do not bind well to the o antibodies (provided that the relative amounts of coated antibody and added protein variant have been adjusted such that the protein backbone when added in similar amounts as the protein variant, binds fully to the coated antibodies). After binding has equilibrated, the supernatant is withdrawn and assayed for functionality to determine whether functional protein variants have reduced antibody-binding capacity. In this mode, the analysis of 5 antibody binding and protein variant functionality are combined to give a single read-out. Optionally, the conditions can be adjusted to lower the affinity between antibody and protein backbone in order to allow protein variants with moderately reduced antibody binding capacity to be free in solution. The affinity can be lowered for instance by adjusting the salt concentration and or pH or by carrying out the incubation in the presence of surface active 0 ingredients or in the presence of competitive modified protein backbone, which has been inactivated (chemically or genetically, as described above for proteases) in order to give no signal in the functionality assay.
In another embodiment, the antibody-coated surfaces are incubated with protein variant and labelled competitive antigen in a ratio that ensures that no or minimal amounts of labelled competitor are bound to the surface when incubated with the protein backbone. After incubation, the supernatant is removed and the amount of bound competitor is determined. The advantage of this approach is that a protein variant that has had an epitope eliminated will be less likely to give a false positive signal by binding through a different epitope. If the protein variant has had an epitope eliminated the subset of antibodies that are specific for this epitope will be unoccupied and allow binding of labelled competitor to the surface, even when the labelled competitor is present in far lower concentration than the protein variant.
Determining functionality
A wide variety of protein functionality assays are available in the literature. Especially, those suitable for automated analysis are useful for this invention. Several have been published in the literature such as protease assays (WO99/34011, Genencor International; J.E. Ness, et al, Nature Biotechn., 17, pp. 893-896, 1999), oxidoreductase assays (Cherry et al., Nature Biotechn., 17 , pp. 379-384, 1999, and assays for several other enzymes (WO99/45143, Novo Nordisk).
Those assays that employ soluble substrates can be employed for direct analysis of functionality of immobilized protein variants. Especially the protease assays described in the Materials and Methods section are useful for that aspect of the invention.
Further analysis of selected protein variants
When protein variants have been selected based on the methods described in this invention, it is desirable to confirm their antibody binding capacity, functionality, and immunogenicity using a purified preparation. For that use, a selected clone should be re-isolated by
conventional microbiological techniques and its characteristics tested in the same assay system, then (if results are confirmed) the protein variant of interest should be expressed in larger scale, purified by conventional techniques, and the reduced antibody binding capacity and the functionality should be examined in detail using dose-response curves and e.g. competitive ELISA (C-ELISA).
The potentially reduced allergenicity (which is likely, but not necessarily true for a variant w. low antibody binding) should be tested in in vivo or in vitro model systems: e.g. an in vitro assays for immunogenicity such as assays based on cytokine expression profiles or other proliferation or differentiation responses of epithelial and other cells incl. B-cells and T-cells. Further, animal models for testing allergenicity should be set up to test a limited number of protein variants that show desired characteristics in vitro. Useful animal models include the guinea pig intratracheal model (GPIT) (Ritz, et al. Fund. Appl. Toxicol., 21, pp. 31-37, 1993), mouse subcutaneous (mouse-SC) (WO 98/30682, Novo Nordisk), the rat intratracheal (rat-IT) (WO 96/17929, Novo Nordisk), and the mouse intranasal (MINT) (Robinson et al., Fund. Appl. Toxicol. 34, PP- 15-24, 1996) models.
The immunogenicity of the protein variant is measured in animal tests, wherein the animals are immunised with the protein variant and the immune response is measured. Specifically, it is of interest to determine the allergenicity of the protein variants by repeatedly exposing the animals to the protein variant by the intratracheal route and following the specific IgG and IgE titers. Alternatively, the mouse intranasal (MINT) test can be used to assess the allergenicity of protein variants. By the present invention the allergenicity is reduced at least 10 times as compared to the allergenicity of the parent protein, preferably 50 times reduced, more preferably 100 times.
However, the present inventors have demonstrated that the performance in a competitive ELISA conelates closely to the immunogenic responses measured in animal tests. To obtain a useful reduction of the allergenicity of a protein, the IgE binding capacity of the protein variant must be reduced to at least below 75 %, preferably below 50 % of the IgE binding capacity of the parent protein as measured by the performance in competitive IgE ELISA, given the value for the IgE binding capacity of the parent protein is set to 100 %.
Materials and Methods
Horse Radish Peroxidase labelled anti-rabbit-Ig (Dako, DK, P217; dilution 1 :1000). Rabbit anti-Savinase polyclonal IgG prepared by conventional means. Rat anti-Savinase IgE.
CovaLink NH2 plates (Nunc, Cat# 459439)
Cyanuric chloride (Aldrich), Acetone (Merck),Tween 20 (Merck), Skim Milk powder (Difco), H2SO4 (Merck), OPD: o-phenylene-diamine: (Kementec cat no. 4260), H2O2, 30% (Merck).
Buffers and Solutions:
Carbonate buffer (0.1 M, pH 10): Na2CO3 10.60 g/L. PBS (pH 7.2): NaCl 8.00 g/L; KC1 0.20 g/L; K2HPO4 1.04 g/L; KH2PO4 0.32 g/L.
Washing buffer: PBS, 0.05% (v/v) Tween 20.
Blocking buffer: PBS, 2% (wt/v) Skim Milk powder.
Dilution buffer: PBS, 0.05% (v/v) Tween 20, 0.5% (wt/v) Skim Milk powder.
Succinyl-Alanine-Alanine-Proline-Phenylalanine-para-nitroanilide. (Suc-AAPF-pNP) Sigma no. S-7388, Mw 624.6 g/mole.
Activation of CovaLink plates:
A fresh stock solution of 10 mg/ml cyanuric chloride in acetone is diluted into PBS, while stirring, to a final concentration of 1 mg/ml and immediately aliquoted into CovaLink NH2 plates (100 microliter per well) and incubated for 5 minutes at room temperature. After three washes with PBS, the plates are dryed at 50°C for 30 minutes, sealed with sealing tape, and stored in plastic bags at room temperature for up to 3 weeks.
Immobilization of antibody/competitive antigen:
Activated CovaLink NH2 plates are coated overnight at 4 °C with 100 microliter of the desired protein (5 micro gram/ml) in PBS followed by 30 min incubation with blocking buffer at room temperature and four washes in PBS-tween.
5 Protease activity:
Analysis with Sue-Ala- Ala-Pro-Phe-pNa:
Proteases cleave the bond between the peptide and p-nitroaniline to give a visible yellow colour absorbing at 405 nm. Briefly, 100 mg suc-AAPF-pNa is dissolved into 1 ml dimethyl sulfoxide (DMSO). 100 microliter of this is diluted into 10 ml with Britton and Robinson o buffer, pH 8.3, and used as substrate for the protease. Reaction is detected kinetically in a spectrophotometer.
Analysis with BODIPY-casein:
The supernatant from the culture medium is diluted 200-fold in the reaction mixture, which s contains 5 microg/ml BODIPY FL-casein (Molecular Probes), 1 mM CaCl , and 50 mM Tris- HCL, pH 7,5. After 60 min. incubation at room temperature, the plates are read at 520 nm with excitation at 485 nm using a FLUOstar (BMG Technologies).
o Examples
Example 1 : Capturing antigen.
In this example a protease antigen is captured on immobilized antibodies and the non-captured 5 fraction is assayed for functionality. This is performed on CovaLink NH2 plates coated with rat anti-Savinase IgE.
Protease libraries producing epitopee variants were established. These variants might or might not be His-tagged. Screening is directly on bacterial culture media, using covalink plates 0 coated with mouse anti-rat IgE monoclonal antibodies satured with anti-savinase specific rat
IgE. The amounts of bound wild type antigen was determined with a anti-wild type polyclonal rabbit antiserum.
The results are shown in figure 1
Example 2: Immobilized competitor.
In this example a 'backbone protease' protease inhibitor is immobilized in the wells and incubated with an excess of the protein variant and labelled antibodies. The level of bound antibodies is determined.
25 microliter sample and 25 microliter anti-Savinase antibody (both diluted in PBS-tween with 0.5 % (w/v) skim milk) are added to the coated well and incubated at room temperature (30 min). The supernatant is removed and the wells are washed three times in PBS-tween.
50 microliter HRP-labelled speciec-spcific anti-Ig antibody is added and incubated 30 min, then the wells are wash three times in PBS-tween. Finally, 50 microliter ODP-H2O2-mixture is added and A492 is measured kinetically to determine the level of bound antibodies. Dilutions are adjusted such that the 'backbone protein' gives none or very little level of bound antibody.
A separate sample is analysed for functionality and the two values are compared.
Desired protein variants show a high level of bound antibody and at the same time a level of functionality similar to the 'backbone protein'.
The results of the anti-savinase IgG binding is shown in figure 2.
Example 3: Immobilization of His-tagged proteases.
The DNA sequence encoding the protease Savinase® (Novo Nordisk A/S, Denmark) is translationally fused to a sequence encoding a polyhistidine tag (His6) and libraries of Savinase®-His6 variants are produced and introduced into Bacillus. After standard culturing, a limited number of Savinase® enzymes of each variant (about 10% of what is secreted by Bacillus carrying the wildtype Savinase gene) are immobilized in the wells of Ni-NTA microtiter plates. The unbound fraction including cells and excess Savinase® is removed, and the plate washed once or twice in a buffer containing 5-20 mM Imidazole. The immobilized Savinase can now be assayed for antibody binding capacity directly, or used modified with activated PEG and washed prior to analysis. The His-tagged Savinase® variants are released from the solid support by the addition of 250 mM Imidazole, and aliquots of the supematants from each well are sampled for antibody binding and functionality assays as described in the previous examples.
The results are shown in figure 3.
Claims
1. A method for screening a library of protein variants for functional variants with reduced antibody binding capacity, comprising the steps of:
(i) generating a diversified library of protein variants starting from a relevant protein backbone,
(ii) transforming the library into suitable host cells,
(iii) culturing host cells,
(iv) sampling each cell culture ,
(v) analysing a sample by determining the antibody binding capacity of the variant protein,
(vi) analysing a sample by determining the functionality of the variant protein.
2. The method according to claim 1, wherein the following steps are added between step (ii) and (iii):
(iib) culturing host cells,
(iic) assaying function,
(iid) selecting host cells expressing functional protein variants.
3. The method according to claim 2, wherein the selected cells in step (iid) are picked by a colony-picker.
4. The method according to claims 1-3, wherein the library diversity is located in epitope areas.
5. The method according to claims 1-4, wherein the protein variants are modified by substitution, addition, and/or deletion of amino acid residues suitable for chemical modification of the protein.
5
6. The method according to claims 1-2, wherein the protein variants are modified by introduction of one or more additional post-translational modification site and expressed in a host suitable for the conesponding in vivo post-translational modification.
lo 7. The method according to claim 6, wherein the site is a N-glycosylation site or a phosphorylation site.
8. The method according to claims 1-7, where the diversified library is randomized at one or more individual positions (DNA codons) at the primer level.
15
9. The method according to claim 8, wherein the library is biased towards amino acids that can be chemically modified.
10. The method according to claim 8, wherein the library is biased towards amino acids that 2o conespond to post-translational modification recognition sequences.
11. The method according to claim 10, wherein the amino acids conespond to N-glycosylation sites or phosphorylation sites.
25 12. The method according to claim 8, wherein the library is biased towards sequences that are not predicted to result in formation of new epitopes.
13. The method according to claims 1-7, where the diversified library is randomized by combination of segments of known sequence,
30
14. The method according to claims 1-13, wherein the library is diversified simultaneously at several discrete sites on the three dimensional structure.
15. The method according to claim 14, wherein the library is assayed with specific polyclonal antibodies.
5 16. The method according to claim 14. Wherein the library is assayed for antigen binding in an assay that requires bivalent antigen-antibody interactions.
17. The method according to claims 1-13, wherein the library is diversified at a single site on the three-dimensional structure.
10
18. The method according to claim 17, wherein the library is assayed with a monospecific antibody.
19. The method according to claim 17, wherein the library is assayed with a monoclonal is antibody.
20. The method according to claims 19, wherein the library is diversified at a single epitope area and assayed with a monospecific antibody purified using the conesponding peptide- phage membrane protein fusion.
20
21. The method according to claims 1-20, wherein the cells in step (ii) are dispensed in a multi-compartment device in a dilution such that each compartment contains an average of 0,2-1,0 cells.
25 22. The method according to claims 1-21, wherein the sample of step (iv) is separated from the host cells by a membrane process.
23. The method according to claims 1-22, wherein the sample is analysed by determining the total content of protein variant.
30
24. The method according to claims 1-22, wherein the sample is analysed by exposure to adverse conditions prior to determining the functionality.
25. The method according to claims 1-22, wherein the sample functionality is analysed both prior to and after exposure to adverse conditions.
5 26. The method according to claim 1-13 and 21-25, wherein the antibodies are derived from animals sensitized with the backbone protein of step (i).
27. The method according to claims 26, wherein the antibodies are derived from animals sensitized by intratracheal exposure
10
28. The method according to claim 1-13 and 21-25, wherein the antibodies are derived from human volunteers that are sensitized to the backbone protein of step (i).
29. The method according to claims 26-28, wherein the antibodies are raised against the same is protein, but expressed in a strain different from the host cells, such as to minimize background binding to host cell impurities.
30. The method according to claims 26-29, wherein the antibodies are contained in serum from the animal or human.
20
31. The method according to claim 30, wherein the antibodies are IgG, IgM and or IgE antibodies.
32. The method according to claims 30-31, wherein the antibodies are antigen-specific 25 antibodies.
33 The method according to claim 32, wherein the antibodies are selected for the binding affinity to specific epitopes.
30 34. The method according to claims 30-33, wherein the antibodies are purified by capturing those that bind to impurities of the culture supernatant.
35. The method according to claims 26-29, wherein the antibodies are monoclonal antibodies.
36. The method according to claim 35, wherein the clones are selected for the binding affinity of their conesponding antibodies to specific epitopes.
5
37. The method according to claims 1-36, wherein the antibody binding is determined from a single dilution of the protein variant.
38. The method according to claims 1-37, wherein the functionality to be determined is o enzyme activity.
39. The method according to claims 1-38, wherein the protein variants are bound to a solid phase.
s 40. The method according to claim 39, wherein the solid phase is a dipstick.
41. The method according to claim 40, wherein the immobilised protein variants are transfened from one test solution to another by sequentially immersing the dipstick in the test solutions. 0
42. The method according to claim 41, wherein the test solution(s) is(are) placed in wells, e.g. in 96 well plates.
43. The method according to claim 39, wherein the solid surface is a microtiter well surface. 5
44. The method according to claim 39, wherein the solid surface is the surface of beads.
45. The method according to claims 39-44, wherein the binding capacity of the solid surface is less than the average protein variant content of the sample such that the surface binding o samples a reproducible amount of protein variant for analysis.
46. The method according to claims 39-45, wherein the protein variant is linked to a fusion peptide which mediates binding to the solid phase.
47. The method according to claims 39-46, wherein the protein variant is modified by 5 chemical conjugation prior to steps (v) and (vi).
48. The method according to claim 47, wherein the protein variant is conjugated with activated PEG.
lo 49. The method according to claim 48, wherein the protein is conjugated to activated PEG molecules of molecular weight ranging from 100 to 5000 Da at a ratio of activated polymer to lysines in protein that is greater than 5.
50. The method according to claims 39-49, wherein the test solution of step (v) comprises is antibodies and competitive backbone protein.
51. The method according to claims 39-50, wherein bound antigen is detected using a primary antigen-specific antibody and a labelled secondary antibody specific for the primary antibody.
20 52. The method according to claims 39-50, wherein the bound antigen is detected using a labelled primary antibody.
53. The method according to claims 50, wherein the competitor is added in an amounts equal to the amount of immobilised protein variant, better at amounts that are lOx higher than the
25 amount of immobilised protein variant, but preferentially more than lOOx higher than the amount of immobilised protein variant.
54. The method according to claims 50 or 53, wherein the selected variants show reduced antibody binding in the presence of competitor, at least 5% reduced, preferably more than
30 10% reduced, more prefened above 50% reduced and most preferably more than 75% reduced.
55 The method according to claims 39-54, wherein the protein variant is eluted from the solid phase prior to determining the functionality in step (vi).
56. The method according to claims 1-38, wherein the antibody binding is determined by 5 agglutination of beads or cells coated with antibodies.
57. The method according to claims 1-38, wherein the antibody binding is determined by IgE antibodies bound to the surface of effector cells
lo 58. The method according to claim 39-54, wherein the protein variants are bound reversibly to a solid phase and are eluted prior to analysis in (v) and (vi).
59. The method according to claims 1-38 or claim 58, wherein antibodies are coated on a solid surface and incubated with sample.
15
60. The method according to claim 59, wherein labelled competitive protein is incubated together with sample.
61. The method according to claim 60, wherein the amount of competitor is smaller than the 20 average amount of protein variant in the sample.
62. The method according to claims 60-61, wherein the competitor is labelled chemically or genetically.
25 63 The method according to claims 60-62, wherein bound labelled competitor is determined after removal of the sample.
64. The method according to claims 1-38 or claim 58, wherein antibody binding is used to capture protein variant and the functionality is determined for the non-captured fraction.
30
65. The method according to claim 64, wherein binding is performed at conditions where antigen-antibody affinity is lowered, such that a moderate change in affinity may lead to non- capture.
66. The method according to claims 64-65, wherein binding is performed in the presence of inactivated competitor, such that a small change in affinity may lead to non-capture.
67. The method according to claim 66, wherein competitor is inactivated chemically, such that it will not affect the functionality assay.
68. The method according to claim 66, wherein the competitor is inactivated by protem engineering, such that it will not affect the functionality assay.
Applications Claiming Priority (3)
| Application Number | Priority Date | Filing Date | Title |
|---|---|---|---|
| DK176599 | 1999-12-09 | ||
| DKPA199901765 | 1999-12-09 | ||
| PCT/DK2000/000682 WO2001031989A2 (en) | 1999-12-09 | 2000-12-08 | High throughput screening (hts) assays for protein variants with reduced antibody binding capacity |
Publications (1)
| Publication Number | Publication Date |
|---|---|
| EP1240513A2 true EP1240513A2 (en) | 2002-09-18 |
Family
ID=8107893
Family Applications (1)
| Application Number | Title | Priority Date | Filing Date |
|---|---|---|---|
| EP00981185A Ceased EP1240513A2 (en) | 1999-12-09 | 2000-12-08 | High throughput screening (hts) assays for protein variants with reduced antibody binding capacity |
Country Status (4)
| Country | Link |
|---|---|
| EP (1) | EP1240513A2 (en) |
| AU (1) | AU779195B2 (en) |
| CA (1) | CA2394377A1 (en) |
| WO (1) | WO2001031989A2 (en) |
Families Citing this family (3)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| DE60335640D1 (en) | 2002-10-01 | 2011-02-17 | Novozymes As | POLYPEPTIDES OF THE GH-61 FAMILY |
| JP5059412B2 (en) | 2004-01-06 | 2012-10-24 | ノボザイムス アクティーゼルスカブ | Alicyclobacillus sp. Polypeptide |
| EP1967584B1 (en) | 2005-08-16 | 2011-03-23 | Novozymes A/S | Polypeptides of strain bacillus SP. P203 |
Family Cites Families (3)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| AU743647B2 (en) * | 1998-01-31 | 2002-01-31 | Mt. Sinai School Of Medicine Of New York University | Methods and reagents for decreasing allergic reactions |
| US20020029391A1 (en) * | 1998-04-15 | 2002-03-07 | Claude Geoffrey Davis | Epitope-driven human antibody production and gene expression profiling |
| ES2220114T3 (en) * | 1998-10-30 | 2004-12-01 | Novozymes A/S | VARIANTS OF LITTLE ALLERGEN PROTEINS. |
-
2000
- 2000-12-08 WO PCT/DK2000/000682 patent/WO2001031989A2/en not_active Ceased
- 2000-12-08 AU AU18522/01A patent/AU779195B2/en not_active Ceased
- 2000-12-08 CA CA002394377A patent/CA2394377A1/en not_active Abandoned
- 2000-12-08 EP EP00981185A patent/EP1240513A2/en not_active Ceased
Non-Patent Citations (1)
| Title |
|---|
| See references of WO0131989A2 * |
Also Published As
| Publication number | Publication date |
|---|---|
| WO2001031989A3 (en) | 2001-11-15 |
| WO2001031989A2 (en) | 2001-05-10 |
| CA2394377A1 (en) | 2001-05-10 |
| AU779195B2 (en) | 2005-01-13 |
| AU1852201A (en) | 2001-05-14 |
Similar Documents
| Publication | Publication Date | Title |
|---|---|---|
| US6309842B1 (en) | Use of modified tethers in screening compound libraries | |
| EP0815257B1 (en) | A method of detecting biologically active substances | |
| EP1124843B1 (en) | Low allergenic protein variants | |
| JP2003518377A (en) | High affinity TCR proteins and methods | |
| Martoglio et al. | Snapshots of membrane-translocating proteins | |
| US20020019009A1 (en) | High throughput screening (HTS) assays | |
| US20030219848A1 (en) | Short enzyme donor fragment | |
| US6686164B1 (en) | Low allergenic protein variants | |
| AU779195B2 (en) | High throughput screening (HTS) assays for protein variants with reduced antibody binding capacity | |
| US6747135B1 (en) | Fluorescent dye binding peptides | |
| AU782044B2 (en) | A method for the assessment of allergenicity | |
| CA2388982A1 (en) | Microtiter plate (mtp) based high throughput screening (hts) assays | |
| US20060188888A1 (en) | Method of screening for improved specific activity of enzymes | |
| CA2425380A1 (en) | Transgenic plants | |
| US20030041354A1 (en) | Transgenic plants | |
| WO2000023463A2 (en) | Fluorescent dye binding peptides | |
| WO2005024012A1 (en) | Increased expression of a modified polypeptide | |
| EP1143011B1 (en) | A method of detecting biologically active substances | |
| US20160108107A1 (en) | Deamidated anti-gluten antibody and uses thereof | |
| WO2003046195A1 (en) | Method for generating a site-specific library of variants |
Legal Events
| Date | Code | Title | Description |
|---|---|---|---|
| PUAI | Public reference made under article 153(3) epc to a published international application that has entered the european phase |
Free format text: ORIGINAL CODE: 0009012 |
|
| 17P | Request for examination filed |
Effective date: 20020709 |
|
| AK | Designated contracting states |
Kind code of ref document: A2 Designated state(s): AT BE CH CY DE DK ES FI FR GB GR IE IT LI LU MC NL PT SE TR |
|
| AX | Request for extension of the european patent |
Free format text: AL;LT;LV;MK;RO;SI |
|
| 17Q | First examination report despatched |
Effective date: 20050429 |
|
| STAA | Information on the status of an ep patent application or granted ep patent |
Free format text: STATUS: THE APPLICATION HAS BEEN REFUSED |
|
| 18R | Application refused |
Effective date: 20080715 |