WO2016101045A1 - Improved oxidoreductases for biocatalysis - Google Patents
Improved oxidoreductases for biocatalysis Download PDFInfo
- Publication number
- WO2016101045A1 WO2016101045A1 PCT/AU2015/050847 AU2015050847W WO2016101045A1 WO 2016101045 A1 WO2016101045 A1 WO 2016101045A1 AU 2015050847 W AU2015050847 W AU 2015050847W WO 2016101045 A1 WO2016101045 A1 WO 2016101045A1
- Authority
- WO
- WIPO (PCT)
- Prior art keywords
- protein
- amino acid
- seq
- isolated
- ancestral
- Prior art date
- Legal status (The legal status is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the status listed.)
- Ceased
Links
Classifications
-
- C—CHEMISTRY; METALLURGY
- C12—BIOCHEMISTRY; BEER; SPIRITS; WINE; VINEGAR; MICROBIOLOGY; ENZYMOLOGY; MUTATION OR GENETIC ENGINEERING
- C12N—MICROORGANISMS OR ENZYMES; COMPOSITIONS THEREOF; PROPAGATING, PRESERVING, OR MAINTAINING MICROORGANISMS; MUTATION OR GENETIC ENGINEERING; CULTURE MEDIA
- C12N9/00—Enzymes; Proenzymes; Compositions thereof; Processes for preparing, activating, inhibiting, separating or purifying enzymes
- C12N9/0004—Oxidoreductases (1.)
- C12N9/0071—Oxidoreductases (1.) acting on paired donors with incorporation of molecular oxygen (1.14)
-
- C—CHEMISTRY; METALLURGY
- C07—ORGANIC CHEMISTRY
- C07K—PEPTIDES
- C07K19/00—Hybrid peptides, i.e. peptides covalently bound to nucleic acids, or non-covalently bound protein-protein complexes
Definitions
- THE present invention relates to the use of oxidoreductases for biocatalysis. More particularly, the invention relates to isolated, engineered oxidoreductase proteins with improved characteristics for use in biocatalysis, and a process for developing these proteins.
- Biocatalysis the use of enzymes to perform chemical transformations of industrial relevance, is emerging as an attractive means of addressing bottlenecks in the synthesis and modification of new chemicals.
- the basic principle underpinning this is that enzymes can catalyse reactions with greater chemo-, regio- and stereo-selectivity than can be accomplished by purely chemical means (Walsh 2001).
- biocatalysts have substantial utility in biosensor technology (Ronkainen et al. 2010) and for bioremediation (Wood 2008).
- Oxidoreductases comprise a major proportion of biocatalysts exploited industrially (Straathof et al. 2002).
- P450 oxidoreductases are amongst the most versatile enzymes known, a property which has made them a high- priority target for exploitation as biocatalysts (Guengerich 2002). Furthermore, P450 oxidoreductases can be used in the preparation of small quantities of metabolites for use as experimental standards, and P450 oxidoreductase-based biocatalytic assays have substantial potential for screening drug candidates.
- P450s tend to be specialized to interact in a highly efficient manner with a relatively small set of structurally related substrates. However, efficiency is generally markedly diminished when microbial P450 forms act on unnatural substrates. In multicellular organisms, certain P450s perform key roles in xenobiotic metabolism, including the metabolism of drugs. These P450s show unusual characteristics compared to other more typical enzyme catalysts in that they act on an extraordinarily wide spectrum of substrates (Rendic 2002); such xenobiotic- metabolizing P450s may therefore be useful in a broader range of biocatalytic applications.
- NADPH-cytochrome P450 reductase referred to variously as NPR or CPR
- Process efficiency considerations mean that it is usually desirable to add high concentrations of substrates and generate high concentrations of product in bioreactors (Straathof et al. 2002). Achieving high substrate concentrations may necessitate the use of organic solvents at concentrations that are deleterious to the activity or stability of most P450s (e.g.
- the invention is broadly directed to the production of "engineered ancestral"
- P450 enzymes and/or redox partners suitable for use with P450 enzymes. It is a preferred object of the invention to provide "engineered ancestral" P450 enzymes and/or redox partners suitable for use with P450 enzymes that display or possess one or more desirable properties that are at least partly absent in extant P450 enzymes and/or redox partners or are relatively enhanced compared to extant P450 enzymes and/or redox partners.
- the invention provides an isolated protein comprising the amino acid sequence set forth in SEQ ID NO: l, or an amino acid sequence at least 80% identical to SEQ ID NO: l , wherein residues X1-X219 in SEQ ID NO: l, or in the amino acid sequence at least 80% identical to SEQ ID NO: 1, may be any amino acid.
- one or more of the residues X1-X219 is selected from the respective group of variable amino acids as set forth in Table 9.
- the isolated protein comprises an amino acid sequence set forth in any one of SEQ ID NO:2-41 or SEQ ID NOS: 544-578.
- the isolated protein of this aspect has P450 enzyme activity.
- This aspect also provides fragments, variants and/or derivatives of the isolated protein.
- the isolated protein comprises an amino acid sequence at least 80% identical to any one of SEQ ID NOS:2-41 or SEQ ID NOS:544-578.
- the invention provides an isolated protein comprising the amino acid sequence set forth in SEQ ID NO: 180, or an amino acid sequence at least 80% identical to SEQ ID NO: 180, wherein residues X 1 -X 163 in SEQ ID NO: 180, or in the amino acid sequence at least 80% identical to SEQ ID NO: 180, may be any amino acid.
- one or more of the residues Xi-Xi 63 is selected from the respective group of variable amino acids set forth in Table 10.
- the isolated protein comprises an amino acid sequence set forth in any one of SEQ ID NOS: 181-250.
- the isolated protein of this aspect has P450 enzyme activity.
- This aspect also provides fragments, variants and/or derivatives of the isolated protein.
- One embodiment provides an isolated protein comprising an amino acid sequence at least 80% identical to the amino acid sequence of any one of SEQ ID NOS: 181-250.
- the invention provides an isolated protein comprising the amino acid sequence set forth in SEQ ID NO:321, or an amino acid sequence at least 80% identical to SEQ ID NOS:321, wherein residues X 1 -X 1 1 in SEQ ID NO:321, or in the amino acid sequence at least 80% identical to SEQ ID NO:321, may be any amino acid.
- one or more of the residues X 1 -X 1 1 is selected from the respective group of variable amino acids set forth in Table 1 1.
- the isolated protein comprises an amino acid set forth in SEQ ID NOS:322-431.
- the isolated protein of this aspect has P450 reductase enzyme activity.
- This aspect also provides fragments, variants and/or derivatives of the isolated protein.
- One embodiment provides an isolated protein comprising an amino acid sequence at least 80% identical to the amino acid sequence of any one of SEQ ID NOS:322-431.
- the invention provides a method of producing or constructing an isolated protein, said method including the step of producing or constructing an engineered ancestral amino acid sequence of at least a fragment of a P450 protein or P450 reductase protein from one or more P450 protein or P450 reductase amino acid sequences that are different to the engineered ancestral amino acid sequence.
- the isolated protein having P450 enzyme activity is the isolated protein of the first aspect.
- the isolated protein comprises an amino acid sequence set forth in SEQ ID NO: l or SEQ ID NO:2.
- Non-limiting examples of the one or more P450 protein amino acid sequences that are different to said engineered ancestral amino acid sequence are set forth in SEQ ID NOS:42-179.
- the isolated protein having P450 enzyme activity is the isolated protein of the second aspect.
- the isolated protein comprises an amino acid sequence set forth in SEQ ID NO: 180 or SEQ ID NO: 181.
- Non-limiting examples of the one or more P450 protein amino acid sequences that are different to said engineered ancestral amino acid sequence are set forth in SEQ ID NOS:251-320.
- the isolated protein having P450 reductase enzyme activity is the isolated protein of the third aspect.
- the isolated protein comprises an amino acid sequence set forth in SEQ ID NO:321 or SEQ ID NO:322.
- Non-limiting examples of the one or more P450 reductase protein amino acid sequences that are different to said engineered ancestral amino acid sequence are set forth in SEQ ID NOS:432-540.
- the invention provides a method of producing or constructing a modified engineered ancestral P450 protein or P450 reductase protein, said method including the step of introducing one or more amino acid substitutions in an amino acid sequence of the engineered ancestral P450 protein or P450 reductase protein to thereby produce or construct the modified engineered ancestral P450 protein or P450 reductase protein.
- said amino acid substitutions are non- conservative amino acid substitutions.
- the modified engineered ancestral P450 protein or P450 reductase protein displays or possesses one or more increased or enhanced properties compared to the engineered ancestral P450 protein or P450 reductase protein.
- the invention provides a modified engineered ancestral
- P450 protein or P450 reductase protein produced according to the method of the sixth aspect.
- the modified P450 enzyme comprises an amino acid sequence set forth in SEQ ID NO: l, any one of SEQ ID NOS:3-41 or SEQ ID NOS:544-578.
- the isolated protein comprises an amino acid sequence at least 80% identical to any one of SEQ ID NOS:3-41 or SEQ ID NOS:544- 578.
- the modified P450 enzyme comprises an amino acid sequence set forth in SEQ ID NO: 180 or any one of SEQ ID NOS: 182-250.
- the isolated protein comprises an amino acid sequence at least 80% identical to any one of SEQ ID NOS: 182-250.
- the modified P450 reductase enzyme comprises an amino acid sequence set forth in SEQ ID NO:321 or any one of SEQ ID NOS:323-431.
- the isolated protein comprises an amino acid sequence at least 80% identical to any one of SEQ ID NOS:323-431.
- the invention provides an isolated nucleic acid encoding an isolated protein, fragment or derivative of any one of the aforementioned aspects, or produced according to the fourth aspect.
- the invention provides a genetic construct comprising an isolated nucleic acid of the eighth aspect.
- the invention provides a host cell comprising the genetic construct of the ninth aspect.
- the invention provides an antibody or antibody fragment which binds and/or is raised against an isolated protein of any one of the aforementioned aspects.
- the antibody or antibody fragment shows at least partial specificity for an isolated protein comprising any one of the amino acid sequences set forth in SEQ ID NOS:2-41 or SEQ ID NOS:544-578.
- the antibody or antibody fragment shows at least partial specificity for an isolated protein comprising any one of the amino acid sequences set forth in SEQ ID NOS: 181- 250.
- the antibody or antibody fragment shows at least partial specificity for an isolated protein comprising any one of the amino acid sequences set forth in 322-431.
- a twelfth aspect of the invention relates to a composition for performing a chemical reaction, said composition comprising one or more isolated proteins according to the aforementioned aspects and one or more buffers, solvents and/or other reagents suitable for performing the chemical reaction.
- a thirteenth aspect of the invention relates to a method of performing a chemical reaction, said method including the step of exposing one or more substrate molecules to one or more isolated proteins according to the aforementioned aspects to thereby perform the chemical reaction.
- a fourteenth aspect of the invention relates to a reaction product produced according to the method of the thirteenth aspect.
- a fifteenth aspect of the invention relates to use of the isolated protein or composition of the aforementioned aspects for biocatalysis, preferably for structure- activity relationship analysis; pharmacological testing; bioremediation; or biosensor technology.
- indefinite articles “a” and “an” are not to be read as singular indefinite articles or as otherwise excluding more than one or more than a single subject to which the indefinite article refers.
- a protein includes one protein, one or more proteins or a plurality of proteins.
- Figure 1 sets out the amino acid sequences SEQ ID NOS: l and 2.
- Figure 2 sets out the amino acid sequences SEQ ID NOS: 180 and 181.
- Figure 3 sets out the amino acid sequences SEQ ID NOS:321 and 322.
- Figure 4 is a schematic representation of the construction of the bicistronic expression vector for CYP3_N1, pCW/3_NlHis/hNPR.
- Figure 5 is a schematic representation of the construction of the monocistronic expression vector for CYP3_N1, pCW/3_NlHis.
- Figure 6 is a schematic representation of the construction of the bicistronic vector for coexpression of the CYP3 N1-EGFP fusion with hCPR, pCW/3_Nl-EYFP His/hNPR.
- Figure 7 is a schematic representation of the construction of the bicistronic expression vector for CYP2D_N1.
- Figure 8 sets out the N-terminal sequences of two inferred ancestors (CYP2D_Nl-nat; and CYP2D N1-FL which is herein referred to as CYP2D N1) and the extant CYP2D proteins.
- the modifications made include changing the second residue to alanine (Gillam et al. 1995, Arch. Biochem. Biophys. 319, 540-550) and using the MAKKTSSKGK leader sequence (von Wachenfeldt et al. 1997, Arch. Biochem. Biophys. 339, 107-114) before the conserved PPGP motif.
- the CYP2D6 modifications used previously have also been included (Gillam et al. 1995, Arch. Biochem. Biophys. 319, 540-550).
- the "FL" sequences are full-length and the "trunc" sequences truncated.
- Figure 9 is a schematic representation of the construction of the bicistronic expression vector for CPR Nl .
- Figure 10 sets out thermostability profiles of extant (CYP3A4, CYP3A5, CYP3A27, and CYP3A37, labelled as 3A4, 3A5, 3A27, and 3A37, respectively) and engineered ancestral (CYP3_N1, labelled as 3_N1) CYP3 proteins, as measured by the percentage of folded protein after treatment with various temperatures.
- Heat treatment comprised heating the protein at the indicated temperature for 60 min, followed by cooling at 4°C and equilibration to room temperature for 5 min.
- FIG 11 sets out thermostability profiles of extant (2D22 from mouse) and ancestor (2D N1) CYP2Ds based on percentage of folded protein after heat treatment.
- 2D N1 refers to the engineered ancestral CYP2D N1 protein, as described herein.
- Heat treatment comprised heating the protein at the indicated temperature for 60 min, followed by cooling at 4°C and equilibration to room temperature for 5 min.
- Figure 12 sets out an Arrhenius plot of extant (CYP3A4, CYP3A5, CYP3A27, and CYP3A37, labelled as 3A4, 3A5, 3A27, and 3A37, respectively) and engineered ancestral (CYP3 N1, labelled as 3_N1) CYP3 proteins based on the percentage of folded protein after heat treatment. Samples were periodically taken and the remaining folded protein was subsequently measured. Heat treatment comprised heating the protein at the indicated temperature for 60 min, followed by cooling at 4°C and equilibration to room temperature for 5 min.
- Figure 13 sets out thermostability profiles of CYP3 N1 expressed in mono- (labelled as 3_N1) and bicistronic (labelled as 3_Nl_hNPR) format based on percentage of folded protein after heat treatment. Heat treatment comprised heating the protein at the indicated temperature for 60 min, followed by cooling at 4°C and equilibration to room temperature for 5 min.
- FIG 14 sets out thermostability profiles of engineered ancestral CPR Nl and extant human CPR (labelled as hNPR). Heat treatment comprised heating the protein at the indicated temperature for 60 min, followed by cooling at 4°C and equilibration to room temperature for 5 min. Data are shown for three preparations of CPR Nl derived from separate cultures (CPR.Nl #1, CPR.Nl #2, and CPR.Nl #3) compared to a pooled preparation of hNPR.
- Figure 16 sets outsolvent stability profiles of engineered ancestral CYP3 N1 (labelled as TS) and extant CYP3A4 (labelled as 3A4) in 10% methanol, 10% DMSO, or 10%) acetonitrile.
- Figure 17 is a schematic representation of the construction of the CYP3 N1 library.
- Figure 18 sets out ligand binding data for engineered ancestral CYP3 N1 (labelled as CYP TS) and extant CYP3A4 (labelled as Parent). '-' indicates that data is not presented.
- Figure 19 sets out CYP3 N1 variant protein amino acid sequences SEQ ID NOS:3-41 respectively, in FASTA format in descending order.
- Figure 20 sets out CYP2D N1 variant protein amino acid sequences SEQ ID NOS: 182-250 respectively, in FASTA format in descending order.
- Figure 22 sets out extant animal CYP3 protein amino acid sequences SEQ ID NOS:42-179 respectively, in FASTA format in descending order.
- Figure 23 sets out extant animal CYP2D protein amino acid sequences SEQ ID NOS:251-320 respectively, in FASTA format in descending order.
- Figure 24 sets out extant animal CPR protein amino acid sequences SEQ ID NOS:432-540 respectively, in FASTA format in descending order.
- Figure 26 sets out certain compounds, and the structure thereof, that may act as substrates for P450 enzymes.
- Figure 27 sets out an overview of the activity of CYP3_N1 (labelled as CYP3 TS ) on various substrates.
- Figure 30 sets out CYP3 N1 variant protein amino acid sequences SEQ ID NOS:544-578 respectively, in FASTA format in descending order.
- Figure 31 sets out metabolism of tamoxifen (100 ⁇ ) to its major demethylated metabolite by CYP3 N1 (blue diamond); CYP3A4 (red square); and CYP2D N1 (green triangle).
- hNPR expressed in the absence of any P450 (purple cross) was included as a control.
- Bacterial membranes containing 0.1 ⁇ of the P450 enzymes indicated plus human NADPH-cytochrome P450 reductase were incubated with 100 ⁇ tamoxifen. At 20 and 120 minutes, respectively, reactions were quenched by addition of two volumes of acetonitrile and protein was removed by sedimentation. Reaction extracts were lyophilised then resuspended for analysis by LC-MS. Results are the means +/- standard deviation of three independent replicates. The percent conversion is based on the mass spectrometer response and the ratio of metabolite to parent in that particular sample.
- Figure 32 sets out metabolism of erythromycin (100 ⁇ ) to its major demethylated metabolite by CYP3 N1 (blue diamond); CYP3A4 (red square); and CYP2D_N1 (green triangle).
- h PR expressed in the absence of any P450 (purple cross) was included as a control.
- Bacterial membranes containing 0.1 ⁇ of the P450 enzymes indicated plus human NADPH-cytochrome P450 reductase were incubated with 100 ⁇ erythromycin. At 20 and 120 minutes, respectively, reactions were quenched by addition of two volumes of acetonitrile and protein was removed by sedimentation. Reaction extracts were lyophilised then resuspended for analysis by LC-MS. Results are the means +/- standard deviation of three independent replicates. The percent conversion is based on the mass spectrometer response and the ratio of metabolite to parent in that particular sample.
- Figure 33 sets out metabolism of erythromycin (10 ⁇ ) to its major demethylated metabolite by CYP3 N1 (blue diamond); CYP3A4 (red square); and CYP2D_N1 (green triangle).
- hNPR expressed in the absence of any P450 (purple cross) was included as a control.
- Bacterial membranes containing 0.1 ⁇ of the P450 enzymes indicated plus human NADPH-cytochrome P450 reductase were incubated with 10 ⁇ erythromycin. At 20 and 120 minutes, respectively, reactions were quenched by addition of two volumes of acetonitrile and protein was removed by sedimentation. Reaction extracts were lyophilised then resuspended for analysis by LC-MS. Results are the means +/- standard deviation of three independent replicates. The percent conversion is based on the mass spectrometer response and the ratio of metabolite to parent in that particular sample.
- Figure 34 sets out metabolism of erythromycin (100 ⁇ ) to its minor demethylated metabolite by CYP2D N1 (green triangle).
- hNPR expressed in the absence of any P450 (purple cross) was included as a control.
- Bacterial membranes containing 0.1 ⁇ of the P450 enzymes indicated plus human NADPH-cytochrome P450 reductase were incubated with 100 ⁇ erythromycin. At 20 and 120 minutes, reactions were quenched by addition of two volumes of acetonitrile and protein was removed by sedimentation. Reaction extracts were lyophilised then resuspended for analysis by LC-MS. Results are the means +/- standard deviation of three independent replicates. No activity was observed with CYP3_N1 or CYP3A4. The percent conversion is based on the mass spectrometer response and the ratio of metabolite to parent in that particular sample.
- Figure 35 sets out metabolism of ticlopidine (100 ⁇ ) to its desaturated metabolite by CYP3 N1 (blue diamond); CYP3A4 (red square); and CYP2D N1 (green triangle).
- hNPR expressed in the absence of any P450 (purple cross) was included as a control.
- Bacterial membranes containing 0.1 ⁇ of the P450 enzymes indicated plus human NADPH-cytochrome P450 reductase were incubated with 100 ⁇ ticlopidine. At the times indicated, reactions were quenched by addition of two volumes of acetonitrile and protein was removed by sedimentation. Reaction extracts were lyophilised then resuspended for analysis by LC-MS. Results are the means +/- standard deviation of three independent replicates. The percent conversion is based on the mass spectrometer response and the ratio of metabolite to parent in that particular sample.
- Figure 36 sets out metabolism of ticlopidine (10 ⁇ ) to its desaturated metabolite by CYP3 N1 (blue diamond); CYP3A4 (red square); and CYP2D N1 (green triangle).
- hNPR expressed in the absence of any P450 (purple cross) was included as a control.
- Bacterial membranes containing 0.1 ⁇ of the P450 enzymes indicated plus human NADPH-cytochrome P450 reductase were incubated with 10 ⁇ ticlopidine. At the times indicated, reactions were quenched by addition of two volumes of acetonitrile and protein was removed by sedimentation. Reaction extracts were lyophilised then resuspended for analysis by LC-MS. Results are the means +/- standard deviation of three independent replicates. The percent conversion is based on the mass spectrometer response and the ratio of metabolite to parent in that particular sample.
- Figure 37 sets out metabolism of ticlopidine (100 ⁇ ) to its doubly desaturated metabolite by CYP3 N1 (blue diamond); CYP3A4 (red square); and CYP2D N1 (green triangle).
- hNPR expressed in the absence of any P450 (purple cross) was included as a control.
- Bacterial membranes containing 0.1 ⁇ of the P450 enzymes indicated plus human NADPH-cytochrome P450 reductase were incubated with 100 ⁇ ticlopidine. At the times indicated, reactions were quenched by addition of two volumes of acetonitrile and protein was removed by sedimentation. Reaction extracts were lyophilised then resuspended for analysis by LC-MS. Results are the means +/- standard deviation of three independent replicates. The percent conversion is based on the mass spectrometer response and the ratio of metabolite to parent in that particular sample.
- CYP3 sequence SEQ ID NO:2 Amino acid sequence of engineered ancestral CYP3 sequence CYP3_N1.
- SEQ ID NOS:3-41 Amino acid sequences of CYP3 N1 variants.
- SEQ ID NOS:42-179 Amino acid sequences of extant animal CYP3 sequences.
- SEQ ID NO: 180 Amino acid sequence of generic engineered ancestral
- CYP2D sequence SEQ ID NO: 181 Amino acid sequence of engineered ancestral CYP2D sequence CYP2D_N1.
- SEQ ID NOS: 182-250 Amino acid sequences of CYP2D N1 variants.
- SEQ ID NOS:251-320 Amino acid sequences of extant animal CYP2D sequences.
- SEQ ID NOS:323-431 Amino acid sequences of CPR Nl variants.
- SEQ ID NOS:432-540 Amino acid sequence of extant animal CPR sequences.
- SEQ ID NO:541 Exemplary nucleotide sequence encoding CYP3 N1
- SEQ ID NO:542 Exemplary nucleotide sequence encoding CYP2D N1.
- SEQ ID NO: 543 Exemplary nucleotide sequence encoding CPR Nl .
- SEQ ID NOS: 544-578 Amino acid sequences of further CYP3 N1 variants.
- This invention relates to the design and production of engineered P450 enzymes and/or P450 redox partners that represent enzymes ancestral to extant P450 enzymes and/or P450 redox partners, respectively.
- the invention is at least partly predicated on the surprising discovery that said engineered ancestral enzymes P450 enzymes and/or P450 redox partners may have one or more increased or enhanced properties relative to one or more extant P450 enzymes and/or P450 redox partners.
- an engineered ancestral enzyme of the invention has substantially increased thermal stability. This is a highly surprising discovery because the temperature conditions to which the hypothetical engineered ancestral enzyme may have been exposed are likely to be substantially similar to the temperature conditions to which one or more corresponding extant enzymes are exposed.
- engineered ancestral CYP3 and CYP2D P450 enzymes have been created which are putative or hypothetical ancestors of extant P450 enzymes.
- amino acid residues in the engineered ancestral CYP3 and CYP2D P450 enzymes have been identified that may be modified to confer, modify or remove one or more properties of the P450 enzymes. These include substrate specificity or promiscuity, thermal stability, pH stability and kinetic properties, although without limitation thereto.
- the invention relates to a P450 redox partner suitable for use with a P450 enzymes such as CYP3 and CYP2D, although without limitation thereto.
- the P450 redox partner is a P450 reductase.
- the P450 redox partner is an "engineered ancestral" CPR P450 reductase.
- amino acid residues in the engineered ancestral CPR P450 reductase have been identified that may be modified to confer, modify or remove one or more properties of the P450 reductase enzymes. These include substrate specificity or promiscuity, thermal stability, pH stability and kinetic properties, although without limitation thereto.
- a “peptide” is a protein having no more than fifty (50) amino acids.
- a “polypeptide” is a protein having more than fifty (50) amino acids.
- oxidizing and “oxidation” refer to increasing the oxidation state of an atom, molecule or ion through a loss or transfer of one or more electrons from the atom, molecule or ion.
- oxidation is accompanied by an associated reduction in oxidation state of another atom, molecule or ion through a gain or transfer of one or more electrons by or to the atom, molecule or ion.
- redox changes in oxidation state
- a "P450 enzyme” or a "protein having P450 enzyme activity” is a protein having one or more activities of a P450 enzyme.
- a preferred activity is monooxygenase activity according to the reaction: RH + 0 2 +NAD(P)H + H + ROH +H 2 0 +NAD(P) + where R is a carbon-containing heteroatom.
- Such reactions include hydroxylation at aromatic and aliphatic centres, epoxidation, N-, 0-, S-dealkylation and N-, and S-oxidation, acyl migration, oxidative dehalogenation, ring expansion, contraction and cleavage, C-C bond cleavage, denitrosation of N-nitrosamines, oxidative ester cleavage, aldehyde scissions ⁇ e.g to alkenes and HCOOH), ipso attack on aromatic ring substituents and N- or O- dearylation.
- One or more other activities may include NO synthase-like activity, reductase activity ⁇ e.g.
- CYP3 protein or "CYP3 enzyme” refers to a particular class or family of proteins having P450 enzyme activity.
- CYP2D protein or “CYP2D enzyme” refers to another particular class or family of proteins having P450 enzyme activity.
- a redox partner is required for redox activity of a P450 enzyme.
- a "redox partner” is a protein that facilitates the transfer of electrons from an electron donor molecule to a P450 enzyme.
- the redox partner of a P450 enzyme is a P450 reductase, although without limitation thereto.
- a "P450 reductase” or a "protein having P450 reductase activity” is a protein having one or more activities of a P450 reductase enzyme.
- a preferred activity is the transfer of electrons from an electron donor molecule to a P450 enzyme.
- CPR protein or “CPR enzyme” refers to the "cytochrome P450 reductase” class or family of proteins having P450 reductase activity.
- a CPR enzyme comprises a "flavin adenine dinucleotide" (“FAD”) -binding domain and a "flavin mononucleotide” (“FMN”) -binding domain.
- FAD flavin adenine dinucleotide
- FMN flavin mononucleotide
- a P450 enzyme may have a diverse array of substrates including testosterone, progesterone, midozalam, nifedipine, tamoxifen, cyclosporin A, erythromycin, cyclophosphamide, paracetamol, lignocaine, ethosuximide, codeine, lovastatin, 7-benzyloxy-4-(trifluoromethyl)-coumarin, 7-benzyloxyresorufin, terfenadine, S-omeprazole and benzyloxyluciferin; pesticides such as organochlorine pesticide, an organophosphate pesticide, or a pyrethroid pesticide; solvents such as perchloroethylene (PCE) or trichloroethylene (TCE); and food contaminants such as carbamate or organophosphate pesticide residues, although without limitation thereto.
- PCE perchloroethylene
- TCE trichloroethylene
- one or more of the residues X1-X219 is an amino acid sequence selected from the respective groups consisting of the variable amino acids set forth in Table 9.
- the isolated protein comprising the amino acid sequence set forth in SEQ ID NO: l, and variants thereof, may be referred to as a protein having P450 activity that is "ancestral" to each of the proteins set forth in SEQ ID NOS:42-179.
- the isolated engineered ancestral CYP3 enzyme comprising the sequence set forth in SEQ ID NO: l, and variants thereof, have one or more improved or enhanced properties compared to one or more of the isolated proteins comprising sequences set forth in SEQ ID NOS:42-179.
- CYP3 N1 The particular engineered ancestral CYP3 enzyme variant comprising the amino acid sequence set forth in SEQ ID NO:2, and described in the EXAMPLES, is herein referred as "CYP3 N1". Additionally, the isolated proteins comprising the amino acid sequences set forth in SEQ ID NOS:3-41 and SEQ ID NOS:544-578 are herein referred to as "CYP3_N1 variants”.
- CYP3 N1 variants set forth in SEQ ID NOS:3-41 are variants of CYP3 N1 which represent the CYP3 N1 variant library constructed as set forth in EXAMPLES 9- 10.
- CYP3 N1 variants set forth in SEQ ID NOS: 544-578 are variants of CYP3 N1 which represent nodes of the evolutionary tree calculated for construction of the engineered ancestral CYP3 protein as described in EXAMPLE 1.
- a CYP3 N1 variant comprising an amino acid sequences set forth in SEQ ID NOS:3-41 or SEQ ID NOS:544-578 has one or more improved or enhanced properties compared to CYP3 N1.
- Non-limiting examples of the one or more improved properties of an engineered ancestral CYP3 enzyme as herein described include: thermal stability, stability in solvents (e.g. organic solvents), metabolite production, catalytic versatility, catalytic efficiency (e.g. the efficiency of coupling of product formation to cofactor consumption), substrate specificity (e.g. increased specificity or increased genericity, as desired) and enzyme kinetic properties (e.g increased V ma X , lower K m ), although without limitation thereto.
- the 70 amino acid sequences respectively set forth in SEQ ID NOS:251-320 are the amino acid sequences of certain isolated CYP2D proteins of animals (i.e. "extant” CYP2D proteins).
- SEQ ID NO: 180 The amino acid sequence set forth in SEQ ID NO: 180 is the amino acid sequence of an isolated engineered ancestral CYP2D protein, as set out in FIG. 2.
- SEQ ID NOS: 181-250 are particular amino acid sequences of variants of SEQ ID NO: 180, comprising variations at one or more of the amino acid residues designated X 1 -X 163
- the location of the residues X 1 -X 163 in SEQ ID NO: 180 with respect to SEQ ID NO: 181 is given in table form in Table 10.
- one or more of the residues X 1 -X 163 is an amino acid sequence selected from the respective groups consisting of the variable amino acids set forth in Table 10.
- the isolated protein comprising the amino acid sequence set forth in SEQ ID NO: 180, and variants thereof, may be referred to as a protein having P450 activity that is ancestral to each of the proteins set forth in SEQ ID NOS:251-320.
- the isolated engineered ancestral CYP2D protein comprising the sequence set forth in SEQ ID NO: 180, and variants thereof, have one or more improved or enhanced properties compared to one or more of the isolated proteins comprising sequences set forth in SEQ ID NOS:251-320.
- CYP2D N1 The particular engineered ancestral CYP2D enzyme variant comprising the amino acid sequence set forth in SEQ ID NO: 181 is referred to herein as "CYP2D N1". Additionally, the isolated proteins comprising the amino acid sequences set forth in SEQ ID NOS: 182-250 are herein referred to as "CYP2D_N1 variants”.
- CYP2D_N1 variants set forth in SEQ ID NOS: 182-250 are variants of
- Non-limiting examples of the one or more improved properties of an engineered ancestral CYP2D enzyme as herein described include: thermal stability, stability in solvents (e.g. organic solvents), metabolite production, catalytic versatility, catalytic efficiency (e.g the efficiency of coupling of product formation to cofactor consumption), ligand binding capacity (e.g. increased or decreased strength of binding; and/or increased or decreased specificity of binding, as desired), substrate specificity (e.g. increased specificity or increased genericity, as desired) and enzyme kinetic properties (e.g increased V ma X , lower K m ), although without limitation thereto.
- the 109 amino acid sequences set forth in SEQ ID NOS:432-540 are the amino acid sequences of certain isolated CPR reductase proteins of animals (i.e "extant" CPR reductase proteins).
- SEQ ID NO:321 is the amino acid sequence of an isolated engineered ancestral CPR protein, as set out in FIG. 3.
- SEQ ID NOS:322-431 are particular amino acid sequences of variants of SEQ ID NO:321, comprising variations at one or more of the amino acid residues designated X 1 -X 1 1 .
- the location of the residues X 1 -X 1 1 in SEQ ID NO:321 with respect to SEQ ID NO:322 is given in table form in Table 1 1.
- one or more of the residues X 1 -X 1 1 is an amino acid sequence selected from the respective groups consisting of the variable amino acids set forth in Table 1 1.
- the isolated engineered ancestral P450 reductase protein comprising the sequence set forth in SEQ ID NO: 321, and variants thereof, have one or more improved or enhanced properties compared to one or more of the isolated proteins comprising sequences set forth in SEQ ID NOS:432-540.
- CPR Nl The particular engineered ancestral CPR enzyme variant comprising the amino acid sequence set forth in SEQ ID NO:322 is referred to herein as "CPR Nl". Additionally, the isolated proteins comprising the amino acid sequences set forth in
- SEQ ID NOS:323-431 are herein referred to as "CPR Nl variants”.
- a CPR Nl variant comprising an amino acid sequence set forth in SEQ ID NOS:323-431 has one or more improved or enhanced properties compared to CPR Nl .
- Certain embodiments also relate to fragments of the isolated proteins disclosed herein.
- a protein "fragment” includes an amino acid sequence that constitutes less than 100%, but at least 20%, 30%, 40%, 50%, 60%, 70%, 80% or 90- 99%) of said isolated protein.
- a protein fragment comprises no more than 6, 10, 12, 15, 20, 30, 40, 50, 60, 70, 80, 90, 100, 120, 150, 200, 250, 300, 350 or 400 contiguous amino acids of the isolated protein.
- the protein fragment has one or more activities of a P450 enzyme or P450 reductase enzyme as hereinbefore described.
- the protein fragment does not comprise an N-terminal "membrane anchor" region.
- the N-terminal membrane anchor region comprises the amino acid positions 1-38 set forth in SEQ ID NO:2, as set forth in FIG. 28.
- Certain embodiments also relate to variants of an isolated protein of the invention.
- the protein variant has one or more activities of a P450 or P450 reductase enzyme as hereinbefore described.
- variant proteins of the invention have one or more amino acids deleted or substituted by different amino acids. It is well understood in the art that some amino acids may be substituted or deleted without an expectation of changing the activity of the protein substantially ( ⁇ 'conservative" substitutions). More substantial changes to activity may be made by introducing substitutions or deletions that are less conservative (' 'non-conservative" substitutions).
- protein variants share at least 70% or 75%, preferably at least 80% or 85% or more preferably at least 90%, 91%, 92%, 93%, 94%, 95%, 96%, 97%, 98% or 99%) sequence identity with an amino acid sequence of the isolated protein.
- said protein variants share at least 70% or 75%, preferably at least 80% or 85% or more preferably at least 90%, 91%, 92%, 93%, 94%, 95%, 96%, 97%, 98% or 99% sequence identity with at least one of the amino acid sequences set forth in SEQ ID NOS:2-41 or SEQ ID NOS: 544-578.
- said variants share at least 70% or 75%, preferably at least 80% or 85% or more preferably at least 90%, 91%, 92%, 93%, 94%, 95%), 96%), 97%), 98%) or 99% sequence identity with at least one of the amino acid sequences set forth in SEQ ID NOS: 181-250.
- protein variants share at least 80% or 85% or more preferably at least 90%, 91%, 92%, 93%, 94%, 95%, 96%, 97%, 98% or 99% sequence identity with an amino acid sequence of the isolated protein, excluding the N- terminal membrane anchor region of said protein.
- the N-terminal membrane anchor region comprises the amino acid positions 1-38 of the amino acid sequence set forth in SEQ ID NO:2, as set forth in Figure 28 and Table 9.
- the N-terminal membrane anchor region of CYP2D N1 comprises the amino acid positions 1-36 as set forth in Table 10.
- N-terminal membrane anchor region of a P450 enzyme or P450 reductase enzyme of the invention may be substantially modified without substantially affecting the function of said protein.
- the protein variant comprises a modified N- terminal "membrane anchor" region.
- CYP3 N1 comprises the modified N-terminal sequence MALLLAVFL at amino acid positions 1-9.
- Such a modified N-terminal membrane anchor region may assist with protein expression, although without limitation thereto.
- sequence comparisons are typically performed by comparing sequences over a “comparison window” to identify and compare local regions of sequence similarity.
- a “comparison window” refers to a conceptual segment of typically 6, 9 or 12 contiguous residues that is compared to a reference sequence.
- the comparison window may comprise additions or deletions (i.e., gaps) of about 20% or less as compared to the reference sequence for optimal alignment of the respective sequences.
- Optimal alignment of sequences for aligning a comparison window may be conducted by computerised implementations of algorithms (Geneworks program by Intelligenetics; GAP, BESTFIT, FAST A, and TFASTA in the Wisconsin Genetics Software Package Release 7.0, Genetics Computer Group, 575 Science Drive Madison, WI, USA, incorporated herein by reference) or by inspection and the best alignment (i.e., resulting in the highest percentage homology over the comparison window) generated by any of the various methods selected.
- sequence identity is used herein in its broadest sense to include the number of exact nucleotide or amino acid matches having regard to an appropriate alignment using a standard algorithm, having regard to the extent that sequences are identical over a window of comparison.
- a “percentage of sequence identity” is calculated by comparing two optimally aligned sequences over the window of comparison, determining the number of positions at which the identical nucleic acid base (e.g., A, T, C, G, U) occurs in both sequences to yield the number of matched positions, dividing the number of matched positions by the total number of positions in the window of comparison (i.e., the window size), and multiplying the result by 100 to yield the percentage of sequence identity.
- sequence identity may be understood to mean the "match percentage” calculated by the DNASIS computer program (Version 2.5 for windows; available from Hitachi Software engineering Co., Ltd., South San Francisco, California, USA).
- Certain embodiments also relate to derivatives of an isolated protein of the present invention.
- the derivative protein has one or more activities of a P450 or P450 reductase enzyme as hereinbefore described.
- derivative proteins have been altered, for example by conjugation or complexing with other chemical moieties, by post-translational modification (e.g phosphorylation, acetylation etc), modification of glycosylation (e.g. adding, removing or altering glycosylation) and/or inclusion of additional amino acid sequences as would be understood in the art.
- post-translational modification e.g phosphorylation, acetylation etc
- modification of glycosylation e.g. adding, removing or altering glycosylation
- inclusion of additional amino acid sequences as would be understood in the art.
- Additional amino acid sequences may include fusion partner amino acid sequences which create a fusion protein.
- fusion partner amino acid sequences may assist in detection and/or purification of the isolated fusion protein.
- Non-limiting examples include metal-binding (e.g polyhistidine) fusion partners, maltose binding protein (MBP), Protein A, glutathione S-transferase (GST), fluorescent protein sequences (e.g. GFP), epitope tags such as myc, FLAG and haemagglutinin tags.
- relevant matrices for affinity chromatography include glutathione-, amylose-, and nickel- or cobalt-conjugated resins respectively.
- Many such matrices are available in kit form, such as the QIAexpressTM system (Qiagen) useful with (HIS 6 ) fusion partners and the Pharmacia GST purification system.
- the fusion partners also have protease cleavage sites, such as for
- Factor X a or Thrombin which allow the relevant protease to partially digest the fusion polypeptide of the invention and thereby liberate the recombinant polypeptide of the invention therefrom.
- the liberated polypeptide can then be isolated from the fusion partner by subsequent chromatographic separation.
- derivatives contemplated by the invention include, but are not limited to, modification to amino acid side chains, incorporation of unnatural amino acids and/or their derivatives during peptide, polypeptide or protein synthesis and the use of crosslinkers and other methods which impose conformational constraints on the isolated protein, fragments and variants disclosed herein.
- the invention provides a method of producing or constructing an isolated protein, said method including the step of producing or constructing an engineered ancestral amino acid sequence of at least a fragment of a P450 protein or P450 reductase protein from one or more P450 protein or P450 reductase amino acid sequences that are different to the engineered ancestral amino acid sequence.
- the engineered ancestral P450 protein or P450 reductase protein displays or possesses one or more increased or enhanced properties compared to at least one of the one or more P450 protein or P450 reductase amino acid sequences that are different to the engineered ancestral amino acid sequence.
- Non-limiting examples of the one or more increased or enhanced properties of the protein include thermal stability, stability in solvents (e.g. organic solvents), metabolite production, catalytic versatility, catalytic efficiency (e.g. the efficiency of coupling of product formation to cofactor consumption), ligand binding capacity (e.g. increased or decreased strength of binding; and/or increased or decreased specificity of binding, as desired), substrate specificity (e.g. increased specificity or increased genericity, as desired), enzyme kinetic properties (e.g. increased Vmax, lower K m ), ability to couple to a diversity of proteins (e.g. P450s and/or other electron acceptor proteins), and ability to use a diversity of electron donors (e.g. both NADPH and NADH, or NADH preferentially as an electron donor), although without limitation thereto.
- solvents e.g. organic solvents
- catalytic efficiency e.g. the efficiency of coupling of product formation to cofactor consumption
- ligand binding capacity e.g. increased or decreased strength of binding
- the engineered ancestral protein having P450 or P450 reductase enzyme activity is distinct from any of a plurality of corresponding "extant" enzymes encoded by respective genomes of different organisms, such as those proteins used to construct or produce the engineered ancestral protein.
- a protein comprising an amino acid sequence comprised by the engineered ancestral protein having P450 or P450 reductase enzyme activity may or may not ever have actually existed until its construction according to the method of the invention.
- the step of producing or constructing an engineered ancestral amino acid sequence of at least a fragment of a P450 protein or P450 reductase protein comprises the use of one or more computational methods for the reconstruction of engineered ancestral DNA and/or amino acid sequences.
- computational methods can include “maximum parsimony” methods, “maximum likelihood” methods, and “Bayesian inference” methods.
- said one or more computational methods include a "marginal likelihood” method and/or a "joint likelihood” method.
- the construction of a sequence of a hypothetical engineered ancestral protein includes the use of one or more software tools for the reconstruction of engineered ancestral DNA and/or amino acid sequences.
- Said software tools may include, although without limitation thereto, FastML (Ashkenazy et al. 2012), Phylobayes 3 (Lartillot et al. 2009) and/or one or more of those listed at http://topicpages.ploscompbiol.Org/wiki/Engineeredancestral_reconstruction#Software, incorporated herein by reference.
- "evolutionary intermediate sequences” may be calculated or constructed using one or more software tools in the process of constructing the engineered ancestral sequence, wherein said evolutionary intermediate sequences may or may not have ever existed prior to construction according to the method.
- an evolutionary tree is constructed for the engineered ancestral P450 or P450 reductase protein using one or more of the aforementioned computational methods and/or software tools, and amino acid sequences which represent nodes within the evolutionary tree are identified.
- a "node ' " within an evolutionary tree represents a common ancestor of the descendants that share or are linked by the node.
- the isolated protein having P450 enzyme activity is an isolated engineered ancestral CYP3 protein as hereinbefore described.
- the isolated protein comprises an amino acid sequence set forth in SEQ ID NO: 1 or SEQ ID NO:2.
- Non-limiting examples of the one or more P450 protein amino acid sequences that are different to said engineered ancestral amino acid sequence are set forth in SEQ ID NOS:42-179.
- Non-limiting examples of the one or more P450 protein amino acid sequences that are different to said engineered ancestral amino acid sequence are set forth in SEQ ID NOS:251-320.
- the isolated protein having P450 reductase enzyme activity is an isolated engineered ancestral CPR protein, as herein described.
- the isolated protein comprises an amino acid sequence set forth in SEQ ID NO:321 or SEQ ID NO:322.
- Non-limiting examples of the one or more P450 protein amino acid sequences that are different to said engineered ancestral amino acid sequence are set forth in SEQ ID NOS:432-540.
- a related aspect provides a modified engineered ancestral P450 protein or P450 reductase protein produced according to the method of this aspect.
- the invention provides a method of producing or constructing a modified engineered ancestral P450 protein or P450 reductase protein, said method including the step of introducing one or more amino acid substitutions in an amino acid sequence of the engineered ancestral P450 protein or P450 reductase protein to thereby produce or construct the modified engineered ancestral P450 protein or P450 reductase protein.
- said amino acid substitutions are non-conservative amino acid substitutions.
- Non-limiting examples of the one or more improved or enhanced properties of the protein include thermal stability, stability in solvents (e.g. organic solvents), metabolite production, catalytic versatility, catalytic efficiency (e.g the efficiency of coupling of product formation to cofactor consumption), substrate specificity (e.g. increased specificity or increased genericity, as desired) and enzyme kinetic properties (e.g. increased V ma X , lower K m ).
- said non-conservative amino acid substitutions comprise functional or functionally important amino acids comprised by an engineered ancestral P450 protein or P450 reductase protein.
- a functional or functionally important amino acid is one which is known or predicted to substantially affect a protein property or function, such as hereinbefore described.
- said non-conservative amino acid substitutions comprise amino acid positions wherein there is low prediction confidence within an engineered ancestral P450 protein or P450 reductase protein amino acid sequence constructed using one or more computational methods and/or software tools, or discrepancy between ancestral P450 protein or P450 reductase protein amino acid sequences constructed using one or more computational methods and/or software tools, such as the computational methods and/or software tools as hereinbefore described.
- Said modifications may comprise amino acid positions wherein there is variation amongst a plurality of hypothetical engineered ancestral protein sequences constructed for the P450 or P450 reductase enzyme using different software tools and/or computational methods, although without limitation thereto.
- positions X 1 -X 1 1 of the generic engineered ancestral CPR protein sequence set forth in SEQ ID NO:322 represent amino acid positions at which there is variation amongst a sequence constructed using a 'Joint Maximum Likelihood' and 'Marginal Maximum Likelihood' method, as hereinabove described.
- CPR Nl represents the sequence constructed using the joint likelihood method.
- said non-conservative amino acid substitutions comprise amino acid positions wherein there is variation amongst extant and/or evolutionary intermediate sequences used to construct a hypothetical engineered ancestral protein sequence, and/or sequences which represent nodes within an evolutionary tree constructed for the hypothetical engineered ancestral P450 or P450 reductase protein, as hereinbefore described.
- positions X 1 -X 219 of the generic engineered ancestral CYP3 protein sequence set forth in SEQ ID NO: l represent amino acid positions at which there is variation amongst corresponding extant and/or evolutionary intermediate sequences used to construct the engineered ancestral CYP3 protein, and/or sequences which represent nodes within the evolutionary tree constructed for the engineered ancestral CYP3 protein.
- positions X 1 -X 163 of the generic engineered ancestral CYP2D protein sequence set forth in SEQ ID NO: 180 represent amino acid positions at which there is variation amongst corresponding extant and/or evolutionary intermediate sequences used to construct the engineered ancestral CYP2D protein, and/or sequences which represent nodes within the evolutionary tree constructed for the engineered ancestral CYP2D protein.
- the modified engineered ancestral P450 protein is a modified engineered ancestral CYP2D protein.
- the modified engineered ancestral CYP2D protein comprises an amino acid sequence set forth in SEQ ID NO: 180 or any one of SEQ ID NOS: 182-250.
- the isolated protein comprises an amino acid sequence at least 80% identical to any one of SEQ ID NOS: 182-250, including at least 85% identical, at least 90% identical, and at least 91%, 92%, 93%, 94%, 95%, 96%, 97%, 98%, and 99% identical.
- the modified engineered ancestral P450 reductase protein is a modified engineered ancestral CPR protein.
- the modified engineered ancestral CPR protein comprises an amino acid sequence set forth in SEQ ID NO:321 or any one of SEQ ID NOS:323-431.
- the isolated protein comprises an amino acid sequence at least 80% identical to any one of SEQ ID NOS:323-431, including at least 85% identical, at least 90% identical, and at least 91%, 92%, 93%, 94%, 95%, 96%, 97%, 98%, and 99% identical.
- an isolated engineered ancestral protein of the invention has enhanced thermal stability compared to one or more corresponding extant animal proteins.
- thermal stability is the ability of a protein to resist "thermal deactivation".
- thermal deactivation refers to the loss of a properly folded state of a protein as a result of exposure to heat, and/or the loss of enzyme activity of a protein as a result of exposure to heat.
- thermal stability may be "thermal deactivation energy"
- E a which can be defined as the minimum energy that is required to cause thermal deactivation of a protein.
- E a for CYP3 N1 is 251 kj/mol (Table 3).
- the isolated engineered ancestral CYP3 N1 protein of the invention has enhanced thermal stability compared to one or more of the isolated proteins comprising SEQ ID NOS:42-179.
- isolated CYP3 N1 displayed substantially greater thermal stability than extant animal CYP3 proteins.
- the isolated CYP2D N1 protein of the invention has enhanced thermal stability compared to one or more of the isolated proteins comprising SEQ ID NOS:251-320.
- isolated CYP2D N1 protein displayed substantially greater thermal stability than an extant animal CYP2D protein.
- the isolated CPR Nl protein of the invention has enhanced thermal stability compared to one or more of the isolated proteins comprising SEQ ID NOS:432-540.
- isolated CPR Nl protein displayed substantially greater thermal stability than extant animal CPR proteins.
- an isolated engineered ancestral protein of the invention has enhanced solvent stability compared to one or more of the corresponding extant animal proteins.
- said solvent is an organic solvent including, but not limited to: a polar protic solvent including an alcohol, e.g. ethanol, methanol, and isopropanol; a polar aprotic solvent e.g. dimethyl sulfoxide (DMSO), and acetonitrile; and a non-polar solvent including saturated and unsaturated hydrocarbon molecules and aromatic molecules.
- a polar protic solvent including an alcohol, e.g. ethanol, methanol, and isopropanol
- a polar aprotic solvent e.g. dimethyl sulfoxide (DMSO), and acetonitrile
- DMSO dimethyl sulfoxide
- non-polar solvent including saturated and unsaturated hydrocarbon molecules and aromatic molecules.
- the isolated engineered ancestral CYP3 N1 molecule of the invention has enhanced stability in an organic solvent comprising methanol, compared to one or more of the isolated proteins comprising SEQ ID NOS:42-179.
- CYP3 N1 displayed substantially higher velocity in methanol compared to extant animal CYP3 proteins.
- the isolated engineered ancestral CYP3 N1 molecule of the invention has enhanced stability in an organic solvent comprising acetonitrile, compared to one or more of the isolated proteins comprising SEQ ID NOS:42-179.
- CYP3 N1 displayed substantially increased velocity in solvents comprising various concentrations of acetonitrile, compared to extant animal CYP3 proteins.
- an isolated engineered ancestral protein of the invention has altered binding affinity for one or more ligands as compared to one or more corresponding extant animal proteins.
- K d dissociation constant
- said one or more ligands comprise a macrolide
- a benzodiazepine e.g. 2-keto benzodiazepines such as diazepam, and imidazo benzodiazepines such as midazolam, although without limitation thereto.
- an isolated engineered ancestral protein of the invention demonstrates altered metabolism of chemicals, such as drugs, as compared to one or more corresponding extant animal proteins.
- an isolated engineered ancestral protein of the invention may demonstrate more rapid conversion of a chemical, such as a drug, to a metabolite of that chemical; and/or result in a greater conversion of a chemical, such as a drug, to a metabolite of that chemical at completion (or near completion).
- CYP3 N1 demonstrated more rapid metabolism of tamoxifen to its major demethylated metabolite, and resulted in a greater conversion of tamoxifen to its major demethylated metabolite after 120 minutes.
- CYP3 N1 demonstrated more rapid metabolism of ticlopidine to its desaturated metabolite, and resulted in a greater conversion of ticlopidine to its desaturated metabolite after 120 minutes.
- the activity of an isolated protein of the invention in a reaction comprising a given substrate produces a different metabolite profile, as compared to the activity of one or more extant animal proteins in a corresponding reaction comprising said substrate.
- the different metabolite profile may comprise one or more metabolites that are present or absent, and/or a relative increase or decrease in one or more metabolites.
- Another aspect of the invention provides an isolated nucleic acid that encodes an isolated protein of the invention, inclusive of fragments, variants and derivatives of the isolated protein.
- nucleic acid is exemplified in SEQ ID NO: 541 (FIG. 25).
- nucleic acid is exemplified in SEQ ID NO:543 (FIG. 25).
- nucleic acid designates single-or double-stranded
- DNA and RNA includes genomic DNA and cDNA.
- RNA includes mRNA, RNA, RNAi, siRNA, cRNA and autocatalytic RNA.
- Nucleic acids may also be DNA-RNA hybrids.
- a nucleic acid comprises a nucleotide sequence which typically includes nucleotides that comprise an A, G, C, T or U base. However, nucleotide sequences may include other bases such as inosine, methylycytosine, methylinosine, methyladenosine and/or thiouridine, although without limitation thereto.
- a "polynucleotide” is a nucleic acid having eighty (80) or more contiguous nucleotides, while an “oligonucleotide " has less than eighty (80) contiguous nucleotides.
- a “probe” may be a single or double-stranded oligonucleotide or polynucleotide, suitably labelled for the purpose of detecting complementary sequences in Northern or Southern blotting, for example.
- a “primer” is usually a single- stranded oligonucleotide, preferably having 15-50 contiguous nucleotides, which is capable of annealing to a complementary nucleic acid "template” and being extended in a template-dependent fashion by the action of a DNA polymerase such as Taq polymerase, RNA-dependent DNA polymerase or SequenaseTM.
- a DNA polymerase such as Taq polymerase, RNA-dependent DNA polymerase or SequenaseTM.
- nucleic acid variants share at least 60% or 65%, preferably at least 70% or 75%, more preferably at least 80% or 85%, and even more preferably at least 90% or 95% nucleotide sequence identity with an isolated nucleic acid that encodes one or more of the isolated proteins of the invention.
- nucleic acid variants hybridize to isolated nucleic acids of the invention, under at least low stringency conditions, preferably under at least medium stringency conditions and more preferably under high stringency conditions.
- Hybridize and Hybridization is used herein to denote the pairing of at least partly complementary nucleotide sequences to produce a DNA-DNA, RNA-RNA or DNA-RNA hybrid.
- Hybrid sequences comprising complementary nucleotide sequences occur through base-pairing between complementary purines and pyrimidines as are well known in the art.
- modified purines for example, inosine, methylinosine and methyladenosine
- modified pyrimidines thiouridine and methylcytosine
- Stringency refers to temperature and ionic strength conditions, and presence or absence of certain organic solvents and/or detergents during hybridisation. The higher the stringency, the higher will be the required level of complementarity between hybridizing nucleotide sequences. “High stringency conditions” designates those conditions under which only nucleic acid having a high frequency of complementary bases will hybridize.
- T m of a duplex DNA decreases by about 1°C with every increase of 1% in the number of mismatched bases.
- isolated nucleic acid variants may be produced using a nucleic acid amplification technique.
- Suitable nucleic acid amplification techniques are well known to the skilled addressee, and include polymerase chain reaction (PCR); strand displacement amplification (SDA); rolling circle replication (RCR); nucleic acid sequence-based amplification (NASBA), Q- ⁇ replicase amplification and helicase- dependent amplification, although without limitation thereto.
- an "amplification product” refers to a nucleic acid product generated by nucleic acid amplification.
- nucleic acid amplification techniques may include quantitative and semi-quantitative techniques such as qPCR, real-time PCR and competitive PCR, as are well known in the art.
- isolated nucleic acid variants may be produced using nucleic acid amplification techniques using one or more degenerate primers based on, or derived from, a nucleotide sequence of an isolated nucleic acid disclosed herein.
- the degenerate primer(s) may be designed to anneal to one or more nucleotide sequences of a variant nucleic acid to thereby facilitate amplification of the variant nucleic acid, or a fragment thereof.
- Yet another aspect of the invention provides a genetic construct that comprises an isolated nucleic acid or variant as herein described and one or more additional nucleotide sequences.
- the genetic construct may be in the form of, or comprise genetic components of, a plasmid, bacteriophage, a cosmid, or a yeast or bacterial artificial chromosome as are well understood in the art.
- Genetic constructs may be suitable for maintenance and propagation of the isolated nucleic acid in bacteria or other host cells, for manipulation by recombinant
- the genetic construct is an expression construct.
- the expression construct comprises one or more nucleic acid or variants disclosed herein operably linked to one or more additional sequences in an expression vector.
- An "expression vector” may be either a self-replicating extra-chromosomal vector such as a plasmid, or a vector that integrates into a host genome.
- operably linked is meant that said additional nucleotide sequence(s) is/are positioned relative to the nucleic acid of the invention preferably to initiate, regulate or otherwise control transcription.
- the additional nucleotide sequences are regulatory sequences. Regulatory nucleotide sequences will generally be appropriate for the host cell used for expression. Numerous types of appropriate expression vectors and suitable regulatory sequences are known in the art for a variety of host cells.
- said one or more regulatory nucleotide sequences may include, but are not limited to, promoter sequences, leader or signal sequences, ribosomal binding sites, transcriptional start and termination sequences, translational start and termination sequences, and enhancer or activator sequences.
- promoters may be either naturally occurring promoters, or hybrid promoters that combine elements of more than one promoter.
- the additional nucleotide sequence is a selectable marker gene to allow the selection of transformed host cells.
- Selectable marker genes are well known in the art and will vary with the host cell used.
- the expression construct may also include an additional nucleotide sequence encoding a fusion partner (typically provided by the expression vector) so that the recombinant polypeptide of the invention is expressed as a fusion protein, as hereinbefore described.
- Isolated proteins of the invention may be prepared by any suitable procedure known to those of skill in the art.
- the isolated protein is a recombinant protein.
- a recombinant isolated protein of the invention may be produced by a method including the steps of:
- Suitable host cells for expression may be prokaryotic or eukaryotic.
- suitable host cells may be mammalian cells, plant cells, yeast cells, insect cells or bacterial cells.
- One preferred host cell for expression of an isolated protein according to the invention is a bacterium.
- the recombinant protein may be conveniently prepared by a person skilled in the art using standard protocols as for example described in Sambrook, et al, MOLECULAR CLONING. A Laboratory Manual (Cold Spring Harbor Press, 1989), in particular Sections 16 and 17; CURRENT PROTOCOLS IN MOLECULAR BIOLOGY Eds. Ausubel et al, (John Wiley & Sons, Inc. 1995-2009), in particular Chapters 10 and 16; and CURRENT PROTOCOLS IN PROTEIN SCIENCE Eds. Coligan et al, (John Wiley & Sons, Inc. 1995-2009), in particular Chapters 1, 5 and 6.
- Another aspect of the invention provides an antibody or antibody fragment which binds, or has been raised against, an isolated protein disclosed herein.
- the antibody or antibody fragment shows at least partial specificity for an isolated protein comprising any one of the amino acid sequences set forth in SEQ ID NOS: 182-250.
- said antibody or antibody fragment does not bind, or demonstrates substantially reduced binding against, one or more of the isolated proteins comprising the amino acid sequences SEQ ID NOS :251-320.
- the antibody or antibody fragment shows at least partial specificity for an isolated protein comprising any one of the amino acid sequences set forth in 322-431.
- the antibody or antibody fragment does not bind, or demonstrates substantially reduced binding against, one or more of the isolated proteins comprising the amino acid sequences SEQ ID NOS:432-540.
- an “antibody” is or comprises an immunoglobulin.
- immunoglobulin includes any antigen-binding protein product of a mammalian immunoglobulin gene complex, including immunoglobulin isotypes IgA, IgD, IgM, IgG and IgE and antigen-binding fragments thereof. Included in the term “immunoglobulin” are immunoglobulins that are chimeric or humanised or otherwise comprise altered or variant amino acid residues, sequences and/or glycosylation, whether naturally occurring or produced by human intervention ⁇ e.g. by recombinant DNA technology).
- Antibodies and antibody fragments may be polycolonal or preferably monoclonal.
- Monoclonal antibodies may be produced using the standard method as for example, described in an article by Kohler & Milstein, 1975, Nature 256, 495, or by more recent modifications thereof as for example described in Chapter 2 of Coligan et al, CURRENT PROTOCOLS IN IMMUNOLOGY, by immortalizing spleen or other antibody producing cells derived from a production species which has been inoculated with an isolated protein or a fragment thereof. It will also be appreciated that antibodies may be produced as recombinant synthetic antibodies or antibody fragments for example by expressing a nucleic acid encoding the antibody or antibody fragment in an appropriate host cell.
- Recombinant synthetic antibody or antibody fragment heavy and light chains may be co-expressed from different expression vectors in the same host cell or expressed as a single chain antibody in a host cell.
- Non-limiting examples of recombinant antibody expression and selection techniques are provided in Chapter 17 of Coligan et al, CURRENT PROTOCOLS IN IMMUNOLOGY and Zuberbuhler et al, 2009, Protein Engineering, Design & Selection 22 169.
- the antibody or antibody fragment is labelled.
- the label may be selected from a group including a chromogen, a catalyst, biotin, digoxigenin, an enzyme, a fluorophore, a chemiluminescent molecule, a radioisotope, a drug or other chemotherapeutic agent, a magnetic bead and/or a direct visual label.
- the antibody or antibody fragment may be used for the detection and/or purification of an isolated protein disclosed herein. Methods of use of the isolated protein
- Another aspect of the invention relates to a method of performing a chemical reaction, said method including the step of exposing a molecule to one or more isolated proteins disclosed herein to thereby perform a chemical reaction in the molecule.
- An aspect of the invention also provides a composition suitable for performing the chemical reaction.
- the composition suitably comprises one or more buffers, salts, solvents and/or other reagents that facilitate or allow the reaction to proceed. It will be understood that such pH buffers, salts, solvents and/or other reagents are well known in the art and may be selected according to the particular type of chemical reaction.
- the method includes exposing the molecule to a protein having P450 activity and/or a redox partner for the protein having P450 activity.
- said method includes the step of exposing the molecule to at least one of:
- the redox partner for said CYP3 protein comprises an isolated engineered ancestral CPR protein disclosed herein.
- the redox partner for said CY2D protein comprises an isolated engineered ancestral CPR protein disclosed herein.
- the redox partner for said CYP2D protein comprises any other suitable redox partner(s), which may include, but is not limited to one or more of cytochrome b and cytochrome b5 reductase; and ferredoxin (e.g. adrenodoxin) and ferrodoxin reductase (e.g. adrenodoxin reductase) or flavodoxin reductase.
- ferredoxin e.g. adrenodoxin
- ferrodoxin reductase e.g. adrenodoxin reductase
- flavodoxin reductase flavodoxin reductase
- the protein having P450 activity comprises an isolated engineered ancestral CYP3 protein disclosed herein.
- the protein having P450 activity comprises an isolated engineered ancestral CYP2D protein disclosed herein.
- the protein having P450 activity comprises any other suitable P450 protein, which may include one or more animal P450 proteins, e.g. CYP1, CYP2, CYP3, CYP4, CYP5, CYP6, CYP7, CYP8, CYP9, CYP1 1, CYP12, CYP17, CYP19, CYP20, CYP21, CYP24, CYP26, CYP27, CYP39, CYP46, and CYP51 animal P450 enzyme classes; microbial P450 proteins, e.g.
- plant P450 proteins e.g. CYP51, CYP74, CYP97, CYP710, CYP71 1, CYP727, CYP746 plant P450 enzymes classes, although without limitation thereto.
- Such chemical reactions include redox reactions, hydroxylation at aromatic and aliphatic centres, epoxidation, N-, 0-, S- dealkylation and N-, and S-oxidation, acyl migration, oxidative dehalogenation, ring expansion, contraction and cleavage, C-C bond cleavage, denitrosation of N- nitrosamines, oxidative ester cleavage, aldehyde scissions (e.g to alkenes and HCOOH), ipso attack on aromatic ring substituents and N- or O-deaiylation, reductions of alkyl halides, N-oxides, nitro compounds, inorganic molecules such as S0 2 , Cr(VI) or NO, desaturations (e.g. dehydrogenations), one electron oxidations, isomerizations and/or phospholipase D activity (e.g phosphate ester hydrolysis).
- redox reactions hydroxylation at aromatic and aliphatic centres,
- reaction may have applications for: the production of fine chemicals (e.g. pharmaceuticals, agrichemicals, fragrances, and dyes); gene therapy; bioremediation; biosensors; diagnostics; plant biotechnology; and/or medicinal chemistry (e.g. drug discovery and pharmacological testing).
- fine chemicals e.g. pharmaceuticals, agrichemicals, fragrances, and dyes
- gene therapy e.g. gene therapy
- bioremediation e.g. drug discovery and pharmacological testing
- biosensors e.g. diagnostics
- plant biotechnology e.g. drug discovery and pharmacological testing
- medicinal chemistry e.g. drug discovery and pharmacological testing.
- an isolated protein disclosed herein may be used for structural diversification of molecules present in molecular libraries, such as natural product libraries, synthetic combinatorial libraries, and/or rationally designed structure-based libraries, although without limitation thereto.
- a "lead” compound refers to a chemical compound that has pharmacological or biological activity likely to be useful for a given purpose; for example, for therapeutic and/or industrial application, although without limitation thereto; but may still have suboptimal properties for said purpose.
- an isolated protein described herein may be used to metabolize an environmental pollutant, for example, although without limitation thereto, a hydrocarbon pollutant such as a diesel, a gasoline, or an oil; a pesticide pollutant such as an organochlorine pesticide, an organophosphate pesticide, or a pyrethroid pesticide; and a solvent pollutant such as perchloroethylene (PCE) or trichloroethylene (TCE).
- the isolated protein may be expressed by a microorganism or a plant, thereby allowing said microorganism or plant to metabolize an environmental pollutant, or improving the efficiency with which said microorganism or plant metabolizes an environmental pollutant.
- an isolated protein described herein may be used to detect metabolites in a human blood sample, for example, drugs or drug metabolites; and/or food contaminants, such as carbamate or organophosphate pesticide residues, although without limitation thereto.
- EXAMPLE 1 Ancestral sequence reconstruction of the cytochrome P450 family 3 (CYP3).
- the initial tree was determined by neighbour-joining (BIONJ) (Gascuel 1997). Bootstrapping analysis was performed to evaluate the tree. Prediction of ancestral nodes of the tree (SEQ ID NOS:544-578) was performed using FastML (Pupko et al. 2000). The last common ancestor of CYP3 was identified and designated ' ⁇ .
- a monocistronic expression construct lacking the hNPR open reading frame
- a CYP3_Nl/enhanced yellow fluorescent protein (EYFP) fusion construct retaining the hCPR expression cassette
- a synthetic, double-stranded, oligonucleotide linker containing Xbal, Smal, Blpl, Sacl, Nsil, Hindlll, and Nhel restriction sites was first ligated into the bicistronic construct (pCW/3_NlHis/hNPR) to facilitate the subsequent removal of hNPR via Blpl digestion followed by relegation of the linearized monocistronic vector (FIG. 5).
- the CYP3 N1 EYFP fusion was obtained by ligating the EYFP fragment from pCW/2C19 F L-EYFPHis/hNPR digested with Sail into the bicistronic construct (FIG. 6).
- EXAMPLE 3 Ancestral sequence reconstruction of the CYP2D subfamily (CYP2D) A total of 70 CYP2D sequences from vertebrate species were collected from Uniprot, NCBI and the cytochrome P450-nomenclature homepage (http://dnelson.utmem.edu/Cytochrome P450.html) database. The sequences were aligned using MEGA 6 (Beta 2) (Tamura et al. 2013), using a Multiple Sequence Comparison by Log-Expectation (MUSCLE) alignment with the following parameters: gap open: -2.9, gap extend: -1.01, hydrophobicity multiplier: 1.2. The alignment was fine-tuned manually to improve its reliability at gap positions.
- MEGA 6 Beta 2
- MUSCLE Multiple Sequence Comparison by Log-Expectation
- the MAKKTSSKGK leader sequence von Wachenfeldht et al. 1997; Rowland et al. 2006
- the PCR product was digested with Ndel and Xbal and cloned into the cognate sites of pCW72D22/hNPR to generate the pCW72D_Nltrunc/h PR bicistronic expression vector.
- CPR sequences were obtained from Uniprot, and by BLASTing the Uniprot sequences against the NCBI (Altschul et al. 1990) database to retrieve additional sequences of high homology, then aligned using the CLUSTALW method (Kyoto University Bioinformatics Center; ⁇ http://www.genome.jp/>)
- Anomalous sequences were removed (e.g. obvious sequencing errors, clearly incomplete open reading frames, etc.), realigned and minor allelic variants and sequences that deviate significantly from the family characteristics (e.g. ⁇ 55% sequence identity to all other forms, either globally or over any section of -20 residues) were pruned such that the ultimate alignment included only one sequence per enzyme.
- the remaining 109 sequences were then aligned with the house fly CPR sequence (designated as the outgroup sequence) and the evolutionary tree derived in MEGA (Tamura et al. 2013) by the ML method, with the designated outgroup as the root. (ML has been shown to be more accurate than MP; Gadagkar et al. 2005).
- This tree was then imported along with the alignment into the FastML web server (Ashkenazy et al. 2012) and prediction of ancestral nodes of the tree (SEQ ID NOS: 182-250) and the last common ancestor using joint maximum likelihood and marginal maximum likelihood methods was performed The last common ancestor made using the joint reconstruction method was designated ' ⁇ and used for further experiments.
- the engineered ancestral amino acid sequence was reverse translated using the Geneart codon optimisation algorithm (Thermo Fisher Scientific: Life Technologies 2014) with codon usage optimised for Escherichia coli.
- the propensity for the corresponding mRNA to fold into stable secondary structure was analysed over the N- terminal nucleotide sequence (from -21 to +96 with respect to start codon) using NUPACK (Zadeh et al. 2011) and minimised by iterative silent changes to the nucleotide sequence.
- the final bicistronic expression vector was generated by Gibson Assembly (Zadeh et al. 2011) of this synthetic gene fragment using the backbone of the pCW74Al 1/hNPR vector from which the hNPR insert had been removed by digestion with Xbal and Hindlll.
- the percentage of folded protein was substantially higher for CYP3 N1 (3N_1) after temperature treatment as compared to the percentage of folded protein for extant CYP3 proteins after temperature treatment.
- the percentage folded protein was ⁇ 100% for CYP3N 1, as compared to ⁇ 0% for the extant CYP3 proteins.
- T50 values for folded protein were calculated, as presented in Table 1.
- the half-life of folded protein was estimated by plotting the exponential decay graph of extant and ancestor proteins heated at various temperatures (50, 55, 60, 65, 70, and 75°C); results are set forth in Table 2.
- the CYP3 N1 (3_N1) ancestor protein displayed substantially elevated half life of folded protein as compared to the extant CYP3 proteins. For example, after heating at 55°C, the half life of CYP3_N1 was ⁇ 566.5 minutes, while the half life for the extant CYP3 proteins was between ⁇ 0.5 minutes and - 1.6 minutes.
- CYP3 protein samples were taken at specific time-points and the residual folded protein was quantified.
- CPR Nl experiments cells were subjected to sub-cellular fractionation and membranes were prepared for analysis using a cytochrome c reductase assay; CPR Nl activity was then assessed via the reduction of the surrogate electron acceptor cytochrome c, as previously described by Guengerich (1994) before and after heat treatment as described above for P450s. Results are set forth in FIG. 14; data are shown for three preparations of CPR Nl derived from separate cultures compared to a pooled preparation of extant human CPR (hNPR).
- the percentage of active enzyme was substantially increased for CPR Nl after temperature treatment as compared to the percentage of active enzyme for extant CPR protein after temperature treatment.
- the percent active enzyme after 60 minutes heating at 55°C was ⁇ 50% for CPR Nl, as compared to ⁇ 10% for extant CPR.
- T50 enzyme activity values were obtained for CPR proteins. 60 ⁇ 50 values were as follows: CPR_N1, 45.0 ⁇ 1.0°C; hNPR, 38.5°C.
- EXAMPLE 6 Characterization of enzyme activity for CYP3_N1 and CYP2D_N1
- Initial screens for activity towards fluorogenic or luminogenic or steroid marker substrates were done with intact cells washed and resuspended as above in WCAB.
- membrane fractions were isolated from bacteria expressing either CYP3 N1, CYP2D N1 or an extant protein, with or without hCPR.
- Recombinant enzymes were expressed in 50 ml cultures, harvested and fractionated according to established procedures.
- An overview of the activity of CYP3 N1 on certain substrates is provided in FIG. 27.
- Resorufin O-dealkylation assays were carried out as described in Chang and Waxman (2006). Results are presented in Table 4.
- the cultures were incubated for 48 h at 25°C, with shaking at 180 rpm, and supplemented at 4 h, 20 h, 28 h and 44 h with glucose (2.75 ⁇ moles per addition, in 50 aliquots of a 10 mg/mL stock) and with ammonium hydroxide (5 ⁇ moles per addition as 5 ⁇ L aliquots of a 1 M solution) in order to provide additional carbon and nitrogen sources.
- the cells were removed by centrifugation (3000 x g, 10 min) and 200 ⁇ L of supernatant from each incubation were added to an equal volume of HPLC-grade acetonitrile.
- nifedipine metabolism was performed as described previously using bacterial membranes containing 0.1 ⁇ P450 in 100 mM Tris buffer pH 7.4 and 5 - 200 ⁇ nifedipine added from a methanolic stock such that the final methanol concentration was 1 % v/v. Reactions were quenched after 5 min with 50 ⁇ L tetrahydrofuran containing 100 ⁇ g/mL nordazepam as internal standard then metabolites were extracted by addition of 800 ⁇ L of ethyl acetate and 200 ⁇ _, sodium carbonate pH 10.
- Extracts were desiccated under a gentle stream of nitrogen then resuspended in 60 ⁇ _, of mobile phase of which 25 ⁇ L, were analysed on a on a C18 column (4.6 x 150 mm, 5 ⁇ , Agilent Technologies), eluted isocratically with mobile phase (55% methanol, 45% water, 0.02% triethylamine, pH 5.0) at flow rate of 0.75 ml/min. Results are presented in Table 4.
- CYP3 N1 Bacterial membranes containing 0.1 ⁇ CYP3 N1 (SEQ ID NO:2) coexpressed with human hNPR (human CPR); extant animal CYP3 protein CYP3A4 (Gillam et al. 1993); CYP2D N1 (SEQ ID NO: 181); or hNPR (human CPR) in the absence of any P450 as a control, were incubated with the substrates listed in Table 7 at two concentrations, 10 and ⁇ , except for cyclosporin A, which was used at 2 and 20 ⁇ .
- Incubations were prepared in 250 ⁇ total volume of 0.1 M potassium phosphate buffer pH 7.4 and initiated by the addition of an NADPH-generating system consisting of 10 mM glucose- 6-phosphate, 250 ⁇ NADP+, and 0.5 U/ml glucose-6-phosphate dehydrogenase. Reactions were quenched at 0, 20 and 120 mins by the addition of two volumes of ice- cold acetonitrile. Precipitated protein was removed by centrifugation and samples were lyophilized then resuspended in 50 % v/v acetonitrile in water.
- Mobile phases consisting of ultra-pure water supplemented with formic acid (0.1% v/v; mobile phase A) and pure acetonitrile (mobile phase B) were employed at a flow rate of 0.5 mL min "1 .
- the gradient used was as follows: 0.0-6.0 min (10-70% mobile phase B); 6.0- 6.7 min (70-90% mobile phase B), then a return to the initial mobile phase composition over 0.01 min.
- the MSE analysis was performed with a Waters Synapt HDMS operating in V-mode positive electrospray ionization (ESI) conditions.
- ESI V-mode positive electrospray ionization
- results of metabolism of erythromycin at a concentration of 100 ⁇ to its major demethylated metabolite by CYP3 N1 (blue diamond); CYP3A4 (red square); and CYP2D N1 (green triangle), and control results for hNPR (purple cross), are given in Figure 32.
- results of metabolism of erythromycin at a concentration of 10 ⁇ to its major demethylated metabolite by CYP3 N1 (blue diamond), CYP3A4 (red square), and CYP2D N1 (green triangle), and control results for hNPR (purple cross) are given in Figure 33.
- Testosterone assays were carried out as described earlier. However solvents (methanol, acetonitrile or DMSO) were included in incubation mixtures to the concentrations indicated in the figures. Results are presented in FIG. 15 and FIG. 16. As will be evident from FIG. 15 and FIG. 16, CYP3 N1 (TS) displayed substantially increased velocity in solvents comprising various concentrations of acetonitrile or methanol than extant animal CYP3 protein CYP3A4 (3A4).
- velocity (pmol/min/pmol P450) in 10% methanol was ⁇ 2 for CYP3_N1 and ⁇ 1.5 for CYP3A4; and velocity (pmol/min/pmol P450) in 10% acetonitrile was ⁇ 0.8 for CYP3_N1 and ⁇ 0.25 for CYP3A4 (FIG. 16).
- Spectral binding assays were performed as described elsewhere (Isin et al. 2008) with bacterial membranes containing CYP3 N1 or CYP3A4 to a final concentration of 0.1 ⁇ P450 in lx TES using an OLIS-modified Aminco DW2A spectrophotometer. Binding data were analysed in Prism using the quadratic form of the binding equation (Isin et al. 2006). Results are presented in FIG. 18. In FIG. 18 CYP3 N1 is labelled as CYP3 TS. EXAMPLE 9. Creation of a CYP3_N1 variant library
- Overlap extension PCR was used for constructing a mutant library containing CYP3 N1 variant-encoding nucleic acids with codon substitutions at multiple sites (Williams et al. 2014).
- Fragment F was amplified using pCW/2C19FL His EYFP/hNPR as template). PCR cycling conditions were as follow: an initial "hot start” at 98°C for 1 min; 29 cycles of 98°C for 10s, T ann for 20 s, and 72°C for 10 s; a polishing stage at 72°C for 10 min; and finally, storage at 4°C until use). PCR products were analysed by using agarose gel electrophoresis and purified using the Wizard® SV Gel and PCR clean-up system (Promega, Australia). Product yields were quantified using a Nanodrop spectrophotometer.
- Equal quantities of matched fragments from each PCR (Al and A2, B l and B2, CI and C2, Dl and D2, El and E2) were combined in a primerless reassembly PCR containing polymerase buffer, 200 ⁇ of each dNTP and 0.6 U Phusion® high fidelity DNA polymerase (New England Biolabs, USA) in a total volume of 30 ⁇ . Cycling conditions consisted of an initial hot start at 98°C for 2 min followed by 14 cycles of 98°C for 10s and 72°C for 30 s before storage at 4°C until use.
- PCR mixtures were then supplemented with additional polymerase buffer, dNTP and Phusion® high fidelity DNA polymerase (New England Biolabs, USA) plus each of the two flanking primers to a final concentration of 0.5 ⁇ in a final volume of 50 iL.
- the fragments were then amplified using the following conditions: an initial hot start at 98°C for 2 min; followed by 14 cycles of 98°C for 10s, and 72°C for 30 s; a polishing stage at 72°C for lOmin; then storage at 4°C until use.
- Fragments A, B, C, D, E and F were combined in stepwise manner. Equal quantities of fragment A were first combined with fragment B in a primerless reassembly PCR followed by PCR with flanking primers as above. After gel purification, fragment AB was then combined with fragment C. The process was continued with the addition of fragments D, E and F in sequence, until the full-length open reading frame was obtained. The full-length mutant sequences were then subcloned into the Ndel and Xbal sites of the pCW/2C19 FL His/hNPR plasmid as set forth in FIG. 17.
- Library variants were expressed as fusion proteins with EYFP from a bicistronic expression vector that allowed co-expression of hCPR. This vector was chosen to enable facile screening to eliminate mutants containing frame-shift mutations or premature stop codons in the subsequent library. The general protocol for P450 expression described above was followed but modified for high throughput format.
- Each CYP3 library variant (bicistronic, monocistronic and EYFP fusion in bicistronic format with hCPR) was grown in duplicate from two independent colonies. All media were as described above but starter cultures were set up in 96 well plates, inoculated from single colonies and incubated overnight at 37°C, with shaking at 400 rpm in a 5 mm orbit microplate shaker.
- Expression cultures were set up in 24-well plates by inoculating 20 ⁇ L starter cultures into 1 mL of TB expression medium. Cultures were incubated at 25°C with shaking at 350 rpm for an initial 5h, then recombinant protein expression was induced by adding arabinose (4 mg/mL), ⁇ -aminolevulinic acid (0.5 mM) and isopropyl- ⁇ -thiogalactopyranoside (IPTG) (lmM) as above. The plates were incubated at 25°C with shaking at 350 rpm for a further 43 h. In some cases, plates were sealed with BreathEasy membranes (Diversified Biotech, Boston, USA) to restrict oxygen availability (microaerobic conditions) during P450 expression.
- 26 variants (comprising amino acid sequences set forth in SEQ ID NOS:3-28) showed the best thermostability in initial tests at 72-74°C and the remaining 13 variants (comprising amino acid sequences set forth in SEQ ID NOS:29-41) were chosen to reflect mutants with lower thermostability.
- shuffled CYP1A library shows both structural integrity and functional diversity.
- P450 2C3 is expressed as a soluble dimer in Escherichia coli following modifications of its N-terminus. Arch. Biochem. Biophys. 339, 107-114 (1997).
Landscapes
- Chemical & Material Sciences (AREA)
- Life Sciences & Earth Sciences (AREA)
- Organic Chemistry (AREA)
- Health & Medical Sciences (AREA)
- Genetics & Genomics (AREA)
- Biochemistry (AREA)
- General Health & Medical Sciences (AREA)
- Medicinal Chemistry (AREA)
- Molecular Biology (AREA)
- Bioinformatics & Cheminformatics (AREA)
- Engineering & Computer Science (AREA)
- Wood Science & Technology (AREA)
- Zoology (AREA)
- Microbiology (AREA)
- General Engineering & Computer Science (AREA)
- Biotechnology (AREA)
- Biomedical Technology (AREA)
- Biophysics (AREA)
- Proteomics, Peptides & Aminoacids (AREA)
- Micro-Organisms Or Cultivation Processes Thereof (AREA)
- Enzymes And Modification Thereof (AREA)
- Preparation Of Compounds By Using Micro-Organisms (AREA)
Abstract
Description
Claims
Applications Claiming Priority (2)
| Application Number | Priority Date | Filing Date | Title |
|---|---|---|---|
| AU2014905277 | 2014-12-24 | ||
| AU2014905277A AU2014905277A0 (en) | 2014-12-24 | Improved oxidoreductases for biocatalysis |
Publications (1)
| Publication Number | Publication Date |
|---|---|
| WO2016101045A1 true WO2016101045A1 (en) | 2016-06-30 |
Family
ID=56148805
Family Applications (1)
| Application Number | Title | Priority Date | Filing Date |
|---|---|---|---|
| PCT/AU2015/050847 Ceased WO2016101045A1 (en) | 2014-12-24 | 2015-12-24 | Improved oxidoreductases for biocatalysis |
Country Status (1)
| Country | Link |
|---|---|
| WO (1) | WO2016101045A1 (en) |
Cited By (1)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| EP4685238A3 (en) * | 2017-06-16 | 2026-03-25 | River Stone Biotech, Inc. | Demethylation of reticuline and derivatives thereof with fungal cytochrome p450 |
Citations (1)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| WO2012038508A1 (en) * | 2010-09-23 | 2012-03-29 | Autodisplay Biotech Gmbh | Surface display of polypeptides containing a metal porphyrin or a flavin |
-
2015
- 2015-12-24 WO PCT/AU2015/050847 patent/WO2016101045A1/en not_active Ceased
Patent Citations (1)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| WO2012038508A1 (en) * | 2010-09-23 | 2012-03-29 | Autodisplay Biotech Gmbh | Surface display of polypeptides containing a metal porphyrin or a flavin |
Non-Patent Citations (5)
| Title |
|---|
| DATABASE GenBank 2 May 2011 (2011-05-02), Database accession no. AD060898.1 * |
| DATABASE UniProtKB "H3BGP9_LATCH)'s NADPH--cytochrome P450 reductase", Database accession no. H3BGP9 * |
| GILLIAM E. M. J.: "Engineering Cytochrome P450 Enzymes", CHEMICAL RESEARCH. TOXICOLOGY., vol. 21, 2008, pages 220 - 231 * |
| KABUMOTO H. ET AL.: "Directed Evolution of the Actinomycete Cytochrome p450 MoxA (CYP105) for Enhanced Activity.", BIOSCIENCE BIOTECHNOLOGY BIOCHEMISTRY, vol. 73, no. 9, 2009, pages 1922 - 1927, XP055187656, DOI: doi:10.1271/bbb.90013 * |
| KUMAR S. ET AL.: "Directed Evolution of Mammalian Cytochrome P450 2B1", THE JOURNAL OF BIOLOGICAL CHEMISTRY, vol. 280, no. 20, 20 May 2005 (2005-05-20), pages 19569 - 19575 * |
Cited By (1)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| EP4685238A3 (en) * | 2017-06-16 | 2026-03-25 | River Stone Biotech, Inc. | Demethylation of reticuline and derivatives thereof with fungal cytochrome p450 |
Similar Documents
| Publication | Publication Date | Title |
|---|---|---|
| Burgdorf et al. | The soluble NAD+-reducing [NiFe]-hydrogenase from Ralstonia eutropha H16 consists of six subunits and can be specifically activated by NADPH | |
| Li et al. | Crystal structure of long-chain alkane monooxygenase (LadA) in complex with coenzyme FMN: unveiling the long-chain alkane hydroxylase | |
| Maier et al. | Molecular characterization of the 56-kDa CYP153 from Acinetobacter sp. EB104 | |
| Gao et al. | NADH oxidase from Lactobacillus reuteri: A versatile enzyme for oxidized cofactor regeneration | |
| Lu et al. | CmlI is an N-oxygenase in the biosynthesis of chloramphenicol | |
| Li et al. | A structural and data-driven approach to engineering a plant cytochrome P450 enzyme | |
| Sode et al. | Increasing the thermal stability of the water-soluble pyrroloquinoline quinone glucose dehydrogenase by single amino acid replacement | |
| Khatri et al. | The CYPome of Sorangium cellulosum So ce56 and identification of CYP109D1 as a new fatty acid hydroxylase | |
| Strillinger et al. | Production of halophilic proteins using Haloferax volcanii H1895 in a stirred-tank bioreactor | |
| Baker et al. | Expression, purification, and biochemical characterization of the flavocytochrome P450 CYP505A30 from Myceliophthora thermophila | |
| Khatri et al. | A natural heme‐signature variant of CYP 267A1 from Sorangium cellulosum So ce56 executes diverse ω‐hydroxylation | |
| Bussmann et al. | RosR (Cg1324), a Hydrogen Peroxide-sensitive MarR-type Transcriptional Regulator of Corynebacterium glutamicum*[S] | |
| Gruber et al. | CbbR and RegA regulate cbb operon transcription in Ralstonia eutropha H16 | |
| Groeneveld et al. | Identification of a novel oxygenase capable of regiospecific hydroxylation of D-limonene into (+)-trans-carveol | |
| CN102080069B (en) | A new cytochrome P450 gene, expressed protein and application thereof | |
| US7402419B2 (en) | Phosphite dehydrogenase mutants for nicotinamide cofactor regeneration | |
| Rolf et al. | Cell‐free protein synthesis for the screening of novel azoreductases and their preferred electron donor | |
| Li et al. | A novel unspecific peroxygenase from Agaricus bisporus var. bisporus for biocatalytic oxyfunctionalisation reactions | |
| Fürst et al. | Exploring the biocatalytic potential of a self‐sufficient cytochrome P450 from thermothelomyces thermophila | |
| Wilderman et al. | Functional characterization of cytochromes P450 2B from the desert woodrat Neotoma lepida | |
| Corsini et al. | Expression of the arsenite oxidation regulatory operon in Rhizobium sp. str. NT‐26 is under the control of two promoters that respond to different environmental cues | |
| Ban et al. | Identification of a vitamin D3-specific hydroxylase genes through actinomycetes genome mining | |
| WO2016101045A1 (en) | Improved oxidoreductases for biocatalysis | |
| Zhang et al. | Rational engineering acyltransferase domain of modular polyketide synthase for expanding substrate specificity | |
| Stierle et al. | P450 in C–C coupling of cyclodipeptides with nucleobases |
Legal Events
| Date | Code | Title | Description |
|---|---|---|---|
| 121 | Ep: the epo has been informed by wipo that ep was designated in this application |
Ref document number: 15871368 Country of ref document: EP Kind code of ref document: A1 |
|
| NENP | Non-entry into the national phase |
Ref country code: DE |
|
| WPC | Withdrawal of priority claims after completion of the technical preparations for international publication |
Ref document number: 2014905277 Country of ref document: AU Date of ref document: 20170616 Free format text: WITHDRAWN AFTER TECHNICAL PREPARATION FINISHED |
|
| 122 | Ep: pct application non-entry in european phase |
Ref document number: 15871368 Country of ref document: EP Kind code of ref document: A1 |


























































