EP1224275A2 - Molecules for diagnostics and therapeutics - Google Patents

Molecules for diagnostics and therapeutics

Info

Publication number
EP1224275A2
EP1224275A2 EP00963614A EP00963614A EP1224275A2 EP 1224275 A2 EP1224275 A2 EP 1224275A2 EP 00963614 A EP00963614 A EP 00963614A EP 00963614 A EP00963614 A EP 00963614A EP 1224275 A2 EP1224275 A2 EP 1224275A2
Authority
EP
European Patent Office
Prior art keywords
cell
proteins
polynucleotide
protein
cells
Prior art date
Legal status (The legal status is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the status listed.)
Withdrawn
Application number
EP00963614A
Other languages
German (de)
French (fr)
Inventor
David M. Hodgson
Stephen E Lincoln
Frank D Russo
Peter A. Spiro
Steven C Banville
Shawn R Bratcher;
Gerard F Dufour
Howard J Cohen
Bruce H Rosen
Purvi Shah
Michael S Chalup
Jennifer L Hillman
Anissa Lee Jones
Jimmy Y Yu
Lila B Greenawalt
Scott R Panzer
Ann M Roseberry
Rachel J Wright
Wensheng Chen
Tommy F Liu
Pierre E Yap
Theresa K Stockdreher
Stefan Amshey
Willy T Fong
Current Assignee (The listed assignees may be inaccurate. Google has not performed a legal analysis and makes no representation or warranty as to the accuracy of the list.)
Incyte Corp
Original Assignee
Incyte Genomics Inc
Priority date (The priority date is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the date listed.)
Filing date
Publication date
Application filed by Incyte Genomics Inc filed Critical Incyte Genomics Inc
Priority claimed from PCT/US2000/025643 external-priority patent/WO2001021836A2/en
Publication of EP1224275A2 publication Critical patent/EP1224275A2/en
Withdrawn legal-status Critical Current

Links

Classifications

    • YGENERAL TAGGING OF NEW TECHNOLOGICAL DEVELOPMENTS; GENERAL TAGGING OF CROSS-SECTIONAL TECHNOLOGIES SPANNING OVER SEVERAL SECTIONS OF THE IPC; TECHNICAL SUBJECTS COVERED BY FORMER USPC CROSS-REFERENCE ART COLLECTIONS [XRACs] AND DIGESTS
    • Y02TECHNOLOGIES OR APPLICATIONS FOR MITIGATION OR ADAPTATION AGAINST CLIMATE CHANGE
    • Y02ATECHNOLOGIES FOR ADAPTATION TO CLIMATE CHANGE
    • Y02A50/00TECHNOLOGIES FOR ADAPTATION TO CLIMATE CHANGE in human health protection, e.g. against extreme weather
    • Y02A50/30Against vector-borne diseases, e.g. mosquito-borne, fly-borne, tick-borne or waterborne diseases whose impact is exacerbated by climate change

Definitions

  • the present invention relates to human molecules and to the use of these sequences in the 5 diagnosis, study, prevention, and treatment of diseases associated with, as well as effects of exogenous compounds on, the expression of human molecules.
  • the human genome is comprised of thousands of genes, many encoding gene products that o function in the maintenance and growth of the various cells and tissues in the body. Aberrant expression or mutations in these genes and their products is the cause of, or is associated with, a variety of human diseases such as cancer and other cell proliferative disorders, autoimmune/inflammatory disorders, infections, developmental disorders, endocrine disorders, metabolic disorders, neurological disorders, gastrointestinal disorders, transport disorders, and connective tissue disorders.
  • the 5 identification of these genes and their products is the basis of an ever-expanding effort to find markers for early detection of diseases, and targets for their prevention and treatment. Therefore, these genes and their products are useful as diagnostics and therapeutics.
  • genes may encode, for example, enzyme molecules, molecules associated with growth and development, biochemical pathway molecules, extracellular information transmission molecules, receptor molecules, intracellular signaling molecules, o membrane transport molecules, protein modification and maintenance molecules, nucleic acid synthesis and modification molecules, adhesion molecules, antigen recognition molecules, secreted and extracellular matrix molecules, cytoskeletal molecules, ribosomal molecules, electron transfer associated molecules, transcription factor molecules, chromatin molecules, cell membrane molecules, and organelle associated molecules.
  • cancer represents a type of cell proliferative disorder that affects nearly every tissue in the body.
  • Cell proliferation must be regulated to maintain both the number of cells and their spatial organization. This regulation depends upon the o appropriate expression of proteins which control cell cycle progression in response to extracellular signals such as growth factors and other mitogens, and intracellular cues such as DNA damage or nutrient starvation.
  • Molecules which directly or indirectly modulate cell cycle progression fall into several categories, including growth factors and their receptors, second messenger and signal transduction proteins, oncogene products, tumor-suppressor proteins, and mitosis-promoting factors.
  • Oncogenes are genes generally derived from normal genes that, through abnormal expression or mutation, can effect the transformation of a normal cell to a malignant one (oncogenesis).
  • Oncoproteins, encoded by oncogenes can affect cell proliferation in a variety of ways and include growth factors, growth factor receptors, intracellular signal transducers, nuclear transcription factors, and cell-cycle control proteins.
  • tumor-suppressor genes are involved in inhibiting cell proliferation. Mutations which cause reduced function or loss of function in tumor-suppressor genes result in aberrant cell proliferation and cancer.
  • DNA-based arrays can provide a simple way to explore the expression of a single polymorphic gene or a large number of genes. When the expression of a single gene is explored, DNA-based arrays are employed to detect the expression of specific gene variants. For example, a p53 tumor suppressor gene array is used to determine whether individuals are carrying mutations that predispose them to cancer. A cytochrome p450 gene array is useful to determine whether individuals have one of a number of specific mutations that could result in increased drug metabolism, drug resistance or drug toxicity. DNA-based array technology is especially relevant for the rapid screening of expression of a large number of genes. There is a growing awareness that gene expression is affected in a global fashion.
  • a genetic predisposition, disease or therapeutic treatment may affect, directly or indirectly, the expression of a large number of genes.
  • the interactions may be expected, such as when the genes are part of the same signaling pathway. In other cases, such as when the genes participate in separate signaling pathways, the interactions may be totally unexpected. Therefore, DNA-based arrays can be used to investigate how genetic predisposition, disease, or therapeutic treatment affects the expression of a large number of genes.
  • SEQ ID NO:l encode, for example, human enzyme molecules.
  • oxidoreductases transferases, hydrolases, lyases, isomerases, and ligases.
  • These enzyme classes are each comprised of numerous substrate-specific enzymes having precise and well regulated functions. These enzymes function by facilitating metabolic processes such as glycolysis, the tricarboxylic cycle, and fatty acid metabolism; synthesis or degradation of amino acids, steroids, phospholipids, alcohols, etc.; regulation of cell signalling, proliferation, inflamation, apoptosis, etc., and through catalyzing critical steps in DNA replication and repair, and the process of translation.
  • Oxidoreductases Oxidoreductases
  • oxidoreductase dehydrogenase or reductase activity
  • Potential cofactors include cytochromes, oxygen, disulfide, iron-sulfur proteins, flavin adenine dinucleotide (FAD), and the nicotinamide adenine dinucleotides NAD and NADP (Newsholme, E.A. and A.R. Leech (1983) Biochemistry for the Medical Sciences, John Wiley and Sons, Chichester, U.K., pp. 779-793).
  • Reductase activity catalyzes the transfer of electrons between substrate(s) and cofactor(s) with concurrent oxidation of the cofactor.
  • the reverse dehydrogenase reaction catalyzes the reduction of a cofactor and consequent oxidation of the substrate.
  • Oxidoreductase enzymes are a broad superfamily of proteins that catalyze numerous reactions in all cells of organisms ranging from bacteria to plants to humans. These reactions include metabolism of sugar, certain detoxification reactions in the liver, and the synthesis or degradation of fatty acids, amino acids, glucocorticoids, estrogens, androgens, and prostaglandins.
  • oxidoreductases oxidases
  • reductases dehydrogenases
  • family members often have distinct cellular localizations, including the cytosol, the plasma membrane, mitochondrial inner or outer membrane, and peroxisomes.
  • Short-chain alcohol dehydrogenases are a family of dehydrogenases that only share 15% to 30% sequence identity, with similarity predominantly in the coenzyme binding domain and the substrate binding domain.
  • SCADs are also involved in synthesis and degradation of fatty acids, steroids, and some prostaglandins, and are therefore implicated in a variety of disorders such as lipid storage disease, myopathy, SCAD deficiency, and certain genetic disorders.
  • retinol dehydrogenase is a SCAD-family member (Simon, A. et al. (1995) J. Biol. Chem.
  • retinol dehydrogenase has been linked to hereditary eye diseases such as autosomal recessive childhood-onset severe retinal dystrophy (Simon, A. et al. (1996) Genomics 36:424-430).
  • Propagation of nerve impulses, modulation of cell proliferation and differentiation, induction of the immune response, and tissue homeostasis involve neurotransmitter metabolism (Weiss, B. (1991) Neurotoxicology 12:379-386; Collins, S.M. et al. (1992) Ann. N.Y. Acad. Sci. 664:415-424; Brown, J.K. and H. Imam (1991) J. Inherit. Metab. Dis. 14:436-458). Many pathways of neurotransmitter metabolism require oxidoreductase activity, coupled to reduction or oxidation of a cofactor, such as NAD7NADH (Newsholme, E.A. and A.R.
  • neurotransmitter degradation pathways that utilize NADVNADH-dependent oxidoreductase activity include those of L-DOPA (precursor of dopamine, a neuronal excitatory compound), glycine (an inhibitory neurotransmitter in the brain and spinal cord), histamine (liberated from mast cells during o the inflammatory response), and taurine (an inhibitory neurotransmitter of the brain stem, spinal cord and retina) (Newsholme, supra, pp. 790, 792).
  • L-DOPA precursor of dopamine, a neuronal excitatory compound
  • glycine an inhibitory neurotransmitter in the brain and spinal cord
  • histamine liberated from mast cells during o the inflammatory response
  • taurine an inhibitory neurotransmitter of the brain stem, spinal cord and retina
  • Epigenetic or genetic defects in neurotransmitter metabolic pathways can result in a spectrum of disease states in different tissues including Parkinson disease and inherited myoclonus (McCance, K.L. and S.E. Hu
  • Tetrahydrofolate is a derivatized glutamate molecule that acts as a carrier, providing activated one-carbon units to a wide variety of biosynthetic reactions, including synthesis of purines, pyrimidines, and the amino acid methionine.
  • Tetrahydrofolate is generated by the activity of a holoenzyme complex called tetrahydrofolate synthase, which includes three enzyme activities: tetrahydrofolate dehydrogenase, tetrahydrofolate cyclohydrolase, and tetrahydrofolate synthetase. o
  • tetrahydrofolate dehydrogenase plays an important role in generating building blocks for nucleic and amino acids, crucial to proliferating cells.
  • 3-Hydroxyacyl-CoA dehydrogenase (3HACD) is involved in fatty acid metabolism. It catalyzes the reduction of 3-hydroxyacyl-CoA to 3-oxoacyl-CoA, with concomitant oxidation of NAD to NADH, in the mitochondria and peroxisomes of eukaryotic cells. In peroxisomes, 3HACD 5 and enoyl-CoA hydratase form an enzyme complex called bifunctional enzyme, defects in which are associated with peroxisomal bifunctional enzyme deficiency. This interruption in fatty acid metabolism produces accumulation of very-long chain fatty acids, disrupting development of the brain, bone, and adrenal glands. Infants born with this deficiency typically die within 6 months (Watkins, P.
  • a ⁇ amyloid- ⁇
  • APP amyloid precursor protein
  • 3HACD has been shown to bind the A ⁇ peptide, and is overexpressed in neurons affected in Alzheimer's disease.
  • an antibody against 3HACD can block the toxic effects of A ⁇ in a 5 cell culture model of Alzheimer's disease (Yan, S. et al. (1997) Nature 389:689-695; OMIM, #602057).
  • Steroids such as estrogen, testosterone, corticosterone, and others, are generated from a common precursor, cholesterol, and are interconverted into one another.
  • a wide variety of enzymes act upon cholesterol, including a number of dehydrogenases.
  • Steroid dehydrogenases such as the 5 hydroxysteroid dehydrogenases, are involved in hypertension, fertility, and cancer (Duax, W.L. and D. Ghosh (1997) Steroids 62:95-100).
  • One such dehydrogenase is 3-oxo-5- ⁇ -steroid dehydrogenase (OASD), a microsomal membrane protein highly expressed in prostate and other androgen-responsive tissues.
  • OASD 3-oxo-5- ⁇ -steroid dehydrogenase
  • OASD catalyzes the conversion of testosterone into dihydrotestosterone, which is the most potent androgen.
  • Dihydrotestosterone is essential for the formation of the male phenotype during o embryogenesis, as well as for proper androgen-mediated growth of tissues such as the prostate and male genitalia.
  • a defect in OASD that prevents the conversion of testosterone into dihydrotestosterone leads to a rare form of male pseudohermaphroditis, characterized by defective formation of the external genitalia (Andersson, S. et al. (1991) Nature 354:159-161; Labrie, F. et al. (1992) Endocrinology 131:1571-1573; OMIM #264600).
  • OASD plays a central role in sexual 5 differentiation and androgen physiology.
  • 17 ⁇ -hydroxysteroid dehydrogenase plays an important role in the regulation of the male reproductive hormone, dihydrotestosterone (DHTT).
  • 17 ⁇ HSD6 acts to reduce levels of DHTT by oxidizing a precursor of DHTT, 3 ⁇ -diol, to androsterone which is readily glucuronidated and removed from tissues.
  • 17 ⁇ HSD6 is active with both androgen and estrogen substrates when 0 expressed in embryonic kidney 293 cells. At least five other isozymes of 17 ⁇ HSD have been identified that catalyze oxidation and/or reduction reactions in various tissues with preferences for different steroid substrates (Biswas, M.G. and D.W. Russell (1997) J. Biol. Chem.
  • 17 ⁇ HSDl preferentially reduces estradiol and is abundant in the ovary and placenta.
  • 17 ⁇ HSD2 catalyzes oxidation of androgens and is present in the endometrium and placenta.
  • 5 17 ⁇ HSD3 is exclusively a reductive enzyme in the testis (Geissler, W.M. et al. (1994) Nat. Genet.
  • Oxidoreductases are components of the fatty acid metabolism pathways in mitochondria and peroxisomes.
  • the main beta-oxidation pathway degrades both saturated and unsaturated fatty acids, o while the auxiliary pathway performs additional steps required for the degradation of unsaturated fatty acids.
  • the auxiliary beta-oxidation enzyme 2,4-dienoyl-CoA reductase catalyzes the removal of even-numbered double bonds from unsaturated fatty acids prior to their entry into the main beta- oxidation pathway.
  • the enzyme may also remove odd-numbered double bonds from unsaturated fatty acids (Koivuranta, K.T. et al. (1994) Biochem. J. 304:787-792; Smeland, T.E. et al.
  • 2,4-dienoyl-CoA reductase is located in both mitochondria and peroxisomes. Inherited deficiencies in mitochondrial and peroxisomal beta-oxidation enzymes are associated with severe diseases, some of which manifest themselves soon after birth and lead to death within a few years. Defects in beta-oxidation are associated with Reye's syndrome, Zellweger syndrome, neonatal adrenoleukodystrophy, infantile Refsum's disease, acyl-CoA oxidase deficiency, 5 and bifunctional protein deficiency (Suzuki, Y. et al. (1994) Am. J. Hum.
  • Peroxisomal beta-oxidation is impaired in cancerous tissue. Although neoplastic human breast epithelial cells have the same number of peroxisomes as do normal cells, fatty acyl-CoA oxidase activity is lower than in control tissue (el Bouhtoury, F. et al. (1992) J. Pathol. 0 166:27-35).
  • Human colon carcinomas have fewer peroxisomes than normal colon tissue and have lower fatty-acyl-CoA oxidase and bifunctional enzyme (including enoyl-CoA hydratase) activities than normal tissue (Cable, S. et al. (1992) Virchows Arch. B Cell Pathol. Incl. Mol. Pathol. 62:221- 226).
  • Another important oxidoreductase is isocitrate dehydrogenase, which catalyzes the conversion of isocitrate to a-ketoglutarate, a substrate of the citric acid cycle.
  • Isocitrate dehydrogenase can be 5 either NAD or NADP dependent, and is found in the cytosol, mitochondria, and peroxisomes. Activity of isocitrate dehydrogenase is regulated developmentally, and by hormones, neurotransmitters, and growth factors.
  • HPR Hydroxypyruvate reductase
  • a peroxisomal 2-hydroxyacid dehydrogenase in the glycolate pathway catalyzes the conversion of hydroxypyruvate to glycerate with the oxidation of o both NADH and NADPH.
  • the reverse dehydrogenase reaction reduces NAD + and NADP + .
  • HPR recycles nucleotides and bases back into pathways leading to the synthesis of ATP and GTP. ATP and GTP are used to produce DNA and RNA and to control various aspects of signal transduction and energy metabolism.
  • Inhibitors of purine nucleotide biosynthesis have long been employed as antiproliferative agents to treat cancer and viral diseases. HPR also regulates biochemical synthesis 5 of serine and cellular serine levels available for protein synthesis.
  • the mitochondrial electron transport (or respiratory) chain is a series of oxidoreductase-type enzyme complexes in the mitochondrial membrane that is responsible for the transport of electrons from NADH through a series of redox centers within these complexes to oxygen, and the coupling of this oxidation to the synthesis of ATP (oxidative phosphorylation). ATP then provides the primary o source of energy for driving a cell's many energy-requiring reactions.
  • the key complexes in the respiratory chain are NADH:ubiquinone oxidoreductase (complex I), succinate:ubiquinone oxidoreductase (complex II), cytochrome c r b oxidoreductase (complex III), cytochrome c oxidase (complex IV), and ATP synthase (complex V) (Alberts, B. et al. (1994) Molecular Biology of the Cell, Garland Publishing, Inc., New York NY, pp. 677-678). All of these complexes are located on 5 the inner matrix side of the mitochondrial membrane except complex II, which is on the cytosolic side.
  • Complex II transports electrons generated in the citric acid cycle to the respiratory chain.
  • the electrons generated by oxidation of succinate to fumarate in the citric acid cycle are transferred through electron carriers in complex II to membrane bound ubiquinone (Q).
  • Q membrane bound ubiquinone
  • Transcriptional regulation of these nuclear-encoded genes appears to be the predominant means for controlling the 5 biogenesis of respiratory enzymes. Defects and altered expression of enzymes in the respiratory chain are associated with a variety of disease conditions.
  • 3-hydroxyisobutyrate dehydrogenase important in valine catabolism, catalyzes the NAD-dependent oxidation of 3-hydroxyisobutyrate to methylmalonate semialdehyde within o mitochondria. Elevated levels of 3-hydroxyisobutyrate have been reported in a number of disease states, including ketoacidosis, methylmalonic acidemia, and other disorders associated with deficiencies in methylmalonate semialdehyde dehydrogenase (Rougraff, P.M. et al. (1989) J. Biol. Chem. 264:5899-5903).
  • IVD isovaleryl-CoA-dehydrogenase
  • IVD is involved in leucine metabolism and catalyzes the oxidation of isovaleryl-CoA to 3-methylcrotonyl-CoA.
  • Human IVD is a tetrameric flavoprotein that is encoded in the nucleus and synthesized in the cytosol as a 45 kDa precursor with a mitochondrial import signal sequence.
  • a genetic deficiency caused by a mutation in the gene encoding IVD, results in the condition known as isovaleric acidemia. This mutation results in inefficient mitochondrial 0 import and processing of the IVD precursor (Vockley, J. et al. (1992) J. Biol. Chem. 267:2494-2501). Transferases
  • Transferases are enzymes that catalyze the transfer of molecular groups. The reaction may involve an oxidation, reduction, or cleavage of covalent bonds, and is often specific to a substrate or to particular sites on a type of substrate. Transferases participate in reactions essential to such 5 functions as synthesis and degradation of cell components, regulation of cell functions including cell signaling, cell proliferation, inflamation, apoptosis, secretion and excretion. Transferases are involved in key steps in disease processes involving these functions. Transferases are frequently classified according to the type of group transferred.
  • methyl transferases transfer one- carbon methyl groups
  • amino transferases transfer nitrogenous amino groups
  • similarly o denominated enzymes transfer aldehyde or ketone, acyl, glycosyl, alkyl or aryl, isoprenyl, saccharyl, phosphorous-containing, sulfur-containing, or selenium-containing groups, as well as small enzymatic groups such as Coenzyme A.
  • Acyl transferases include peroxisomal carnitine octanoyl transferase, which is involved in the fatty acid beta-oxidation pathway, and mitochondrial carnitine palmitoyl transferases, involved in 5 fatty acid metabolism and transport. Choline O-acetyl transferase catalyzes the biosynthesis of the neurotransmitter acetylcholine.
  • Amino transferases play key roles in protein synthesis and degradation, and they contribute to other processes as well.
  • the amino transferase 5-aminolevulinic acid synthase catalyzes the addition of succinyl-CoA to glycine, the first step in heme biosynthesis.
  • Other amino transferases 5 participate in pathways important for neurological function and metabolism.
  • glutamine- phenylpyruvate amino transferase also known as glutamine transaminase K (GTK) catalyzes several reactions with a pyridoxal phosphate cofactor.
  • GTK glutamine transaminase K
  • GTK catalyzes the reversible conversion of L- glutamine and phenylpyruvate to 2-oxoglutaramate and L-phenylalanine.
  • Other amino acid substrates for GTK include L-methionine, L-histidine, and L-tyrosine.
  • GTK also catalyzes the conversion of o kynurenine to kynurenic acid, a tryptophan metabolite that is an antagonist of the N-methyl-D- aspartate (NMD A) receptor in the brain and may exert a neuromodulatory function. Alteration of the kynurenine metabolic pathway may be associated with several neurological disorders.
  • GTK also plays a role in the metabolism of halogenated xenobiotics conjugated to glutathione, leading to nephrotoxicity in rats and neurotoxicity in humans.
  • GTK is expressed in kidney, liver, and brain. 5
  • Both human and rat GTKs contain a putative pyridoxal phosphate binding site (ExPASy ENZYME: EC 2.6.1.64; Perry, S.J. et al. (1993) Mol. Pharmacol. 43:660-665; Perry, S. et al. (1995) FEBS Lett. 360:277-280; and Alberati-Giani, D. et al. (1995) J. Neurochem. 64:1448-1455).
  • a second amino transferase associated with this pathway is kynurenine/ ⁇ -aminoadipate amino transferase (AadAT).
  • AadAT catalyzes the reversible conversion of ⁇ -aminoadipate and ⁇ -ketoglutarate to ⁇ -ketoadipate o and L-glutamate during lysine metabolism.
  • AadAT also catalyzes the transamination of kynurenine to kynurenic acid.
  • a cytosolic AadAT is expressed in rat kidney, liver, and brain (Nakatani, Y. et al. (1970) Biochim. Biophys. Acta 198:219-228; Buchli, R. et al. (1995) J. Biol. Chem. 270:29330- 29335).
  • Glycosyl transferases include the mammalian UDP-glucouronosyl transferases, a family of 5 membrane-bound microsomal enzymes catalyzing the transfer of glucouronic acid to lipophilic substrates in reactions that play important roles in detoxification and excretion of drugs, carcinogens, and other foreign substances.
  • Another mammalian glycosyl transferase mammalian UDP-galactose- ceramide galactosyl transferase, catalyzes the transfer of galactose to ceramide in the synthesis of galactocerebrosides in myelin membranes of the nervous system.
  • the UDP-glycosyl transferases o share a conserved signature domain of about 50 amino acid residues (PROSITE: PDOC00359, http://expasy.hcuge.ch/sprot/prosite.html).
  • Methyl transferases are involved in a variety of pharmacologically important processes. Nicotinamide N-methyl transferase catalyzes the N-methylation of nicotinamides and other pyridines, an important step in the cellular handling of drugs and other foreign compounds. 5 Phenylethanolamine N-methyl transferase catalyzes the conversion of noradrenalin to adrenalin. 6-0- methylguanine-DNA methyl transferase reverses DNA methylation, an important step in carcinogenesis.
  • Uroporphyrin-III C-methyl transferase which catalyzes the transfer of two methyl groups from S-adenosyl-L-methionine to uroporphyrinogen III, is the first specific enzyme in the biosynthesis of cobalamin, a dietary enzyme whose uptake is deficient in pernicious anemia.
  • Protein- 5 arginine methyl transferases catalyze the posttranslational methylation of arginine residues in proteins, resulting in the mono- and dimethylation of arginine on the guanidino group.
  • Substrates include histones, myelin basic protein, and heterogeneous nuclear ribonucleoproteins involved in mRNA processing, splicing, and transport.
  • Protein-arginine methyl transferase interacts with proteins upregulated by mitogens, with proteins involved in chronic lymphocytic leukemia, and with 0 interferon, suggesting an important role for methylation in cytokine receptor signaling (Lin, W.-J. et al. (1996) J. Biol. Chem. 271:15034-15044; Abramovich, C. et al. (1997) EMBO J. 16:260-266; and Scott, H.S. et al. (1998) Genomics 48:330-340).
  • Phosphotransferases catalyze the transfer of high-energy phosphate groups and are important in energy-requiring and -releasing reactions.
  • the metabolic enzyme creatine kinase catalyzes the 5 reversible phosphate transfer between creatine/creatine phosphate and ATP/ADP.
  • Glycocyamine kinase catalyzes phosphate transfer from ATP to guanidoacetate
  • arginine kinase catalyzes phosphate transfer from ATP to arginine.
  • a cysteine-containing active site is conserved in this family (PROSITE: PDOC00103).
  • Prenyl transferases are heterodimers, consisting of an alpha and a beta subunit, that catalyze 0 the transfer of an isoprenyl group.
  • An example of a prenyl transferase is the mammalian protein farnesyl transferase.
  • the alpha subunit of farnesyl transferase consists of 5 repeats of 34 amino acids each, with each repeat containing an invariant tryptophan (PROSITE: PDOC00703).
  • Saccharyl transferases are glycating enzymes involved in a variety of metabolic processes. Oligosacchryl transferase-48, for example, is a receptor for advanced glycation endproducts. 5 Accumulation of these endproducts is observed in vascular complications of diabetes, macrovascular disease, renal insufficiency, and Alzheimer's disease (Thornalley, P.J. (1998) Cell Mol. Biol. (Noisy- Le-Grand) 44:1013-1023).
  • Coenzyme A (Co A) transferase catalyzes the transfer of Co A between two carboxylic acids.
  • Succinyl CoA:3-oxoacid CoA transferase for example, transfers CoA from succinyl-CoA to a o recipient such as acetoacetate.
  • Acetoacetate is essential to the metabolism of ketone bodies, which accumulate in tissues affected by metabolic disorders such as diabetes (PROSITE: PDOC00980). Hydrolases
  • Hydrolysis is the breaking of a covalent bond in a substrate by introduction of a molecule of water.
  • the reaction involves a nucleophilic attack by the water molecule's oxygen atom on a target 5 bond in the substrate.
  • the water molecule is split across the target bond, breaking the bond and generating two product molecules.
  • Hydrolases participate in reactions essential to such functions as synthesis and degradation of cell components, and for regulation of cell functions including cell signaling, cell proliferation, inflamation, apoptosis, secretion and excretion. Hydrolases are involved in key steps in disease processes involving these functions.
  • Hydrolytic enzymes may 5 be grouped by substrate specificity into classes including phosphatases, peptidases, lysophospholipases, phosphodiesterases, glycosidases, and glyoxalases.
  • LPLs l o Lysophospholipases
  • LPLs regulate intracellular lipids by catalyzing the hydrolysis of ester bonds to remove an acyl group, a key step in lipid degradation.
  • Small LPL isoforms approximately 15-30 kD, function as hydrolases; larger isoforms function both as hydrolases and transacylases.
  • a particular substrate for LPLs, lysophosphatidylcholine, causes lysis of cell membranes. LPL activity is regulated by signaling molecules important in numerous pathways, including the inflammatory
  • Peptidases also called proteases, cleave peptide bonds that form the backbone of peptide or protein chains. Proteolytic processing is essential to cell growth, differentiation, remodeling, and homeostasis as well as inflammation and immune response. Since typical protein half -lives range from hours to a few days, peptidases are continually cleaving precursor proteins to their active form,
  • Peptidases function in bacterial, parasitic, and viral invasion and replication within a host.
  • peptidases include trypsin and chymotrypsin (components of the complement cascade and the blood-clotting cascade) lysosomal cathepsins, calpains, pepsin, renin, and chymosin (Beynon, R.J. and J.S. Bond (1994) Proteolytic Enzymes: A Practical Approach, Oxford University Press, New
  • Phosphodiesterases catalyze the hydrolysis of one of the two ester bonds in a phosphodiester compound. Phosphodiesterases are therefore crucial to a variety of cellular processes. Phosphodiesterases include DNA and RNA endo- and exo-nucleases, which are essential to cell growth and replication as well as protein synthesis. Another phosphodiesterase is acid
  • sphingomyeUnase which hydrolyzes the membrane phosphoUpid sphingomyelin to ceramide and phosphorylcholine.
  • Phosphorylcholine is used in the synthesis of phosphatidylcholine, which is involved in numerous intracellular signaling pathways.
  • Ceramide is an essential precursor for the generation of gangUosides, membrane lipids found in high concentration in neural tissue. Defective acid sphingomyeUnase phosphodiesterase leads to a build-up of sphingomyelin molecules in
  • Glycosidases catalyze the cleavage of hemiacetyl bonds of glycosides, which are compounds that contain one or more sugar.
  • Mammalian lactase-phlorizin hydrolase for example, is an intestinal enzyme that splits lactose.
  • Mammalian beta-galactosidase removes the terminal galactose from gangliosides, glycoproteins, and glycosaminoglycans, and deficiency of this enzyme is associated 5 with a gangliosidosis known as Morquio disease type B.
  • Vertebrate lysosomal alpha-glucosidase which hydrolyzes glycogen, maltose, and isomaltose
  • vertebrate intestinal sucrase-isomaltase which hydrolyzes sucrose, maltose, and isomaltose
  • the glyoxylase system is involved in gluconeogenesis, the production of glucose from o storage compounds in the body. It consists of glyoxylase I, which catalyzes the formation of S-D- lactoylglutathione from methyglyoxal, a side product of triose-phosphate energy metaboUsm, and glyoxylase II, which hydrolyzes S-D-lactoylglutathione to D-lactic acid and reduced glutathione. Glyoxylases are involved in hyperglycemia, non-insuUn-dependent diabetes mellitus, the detoxification of bacterial toxins, and in the control of cell proliferation and microtubule assembly. 5 Lvases
  • Lyases are a class of enzymes that catalyze the cleavage of C-C, C-O, C-N, C-S, C-(halide), P-0 or other bonds without hydrolysis or oxidation to form two molecules, at least one of which contains a double bond (Stryer, L. (1995) Biochemistry W.H. Freeman and Co. New York, NY p.620). Lyases are critical components of cellular biochemistry with roles in metabolic energy 0 production including fatty acid metabolism, as well as other diverse enzymatic processes. Further classification of lyases reflects the type of bond cleaved as well as the nature of the cleaved group.
  • the group of C-C lyases include carboxyl-lyases (decarboxylases), aldehyde-lyases (aldolases), oxo-acid-lyases and others.
  • the C-O lyase group includes hydro-lyases, lyases acting on polysaccharides and other lyases.
  • the C-N lyase group includes ammonia-lyases, amidine-lyases, 5 amine-lyases (deaminases) and other lyases.
  • lyases Proper regulation of lyases is critical to normal physiology. For example, mutation induced deficiencies in the uroporphyrinogen decarboxylase can lead to photosensitive cutaneous lesions in the genetically-Unked disorder familial porphyria cutanea tarda (Mendez, M. et al. (1998) Am. J. Genet. 63: 1363- 1375). It has also been shown that adenosine deaminase (ADA) deficiency stems o from genetic mutations in the ADA gene, resulting in the disorder severe combined immunodeficiency disease (SCID) (Hershfield, M.S. (1998) Semin. Hematol. 35:291-298). Isomerases
  • Isomerases are a class of enzymes that catalyze geometric or structural changes within a molecule to form a single product. This class includes racemases and epimerases, cis-trans- 5 isomerases, intramolecular oxidoreductases, intramolecular transferases (mutases) and intramolecular lyases. Isomerases are critical components of cellular biochemistry with roles in metabolic energy production including glycolysis, as well as other diverse enzymatic processes (Stryer, L. (1995) Biochemistry, W.H. Freeman and Co., New York NY, pp.483-507).
  • Racemases are a subset of isomerases that catalyze inversion of a molecules configuration 5 around the asymmetric carbon atom in a substrate having a single center of asymmetry, thereby interconverting two racemers.
  • Epimerases are another subset of isomerases that catalyze inversion of configuration around an asymmetric carbon atom in a substrate with more than one center of symmetry, thereby interconverting two epimers. Racemases and epimerases can act on amino acids and derivatives, hydroxy acids and derivatives, as well as carbohydrates and derivatives.
  • the l o interconversion of UDP-galactose and UDP-glucose is catalyzed by UDP-galactose-4' -epimerase.
  • Oxidoreductases can be isomerases as well. Oxidoreductases catalyze the reversible transfer
  • This class of enzymes includes dehydrogenases, hydroxylases, oxidases, oxygenases, peroxidases, and reductases.
  • oxidases Proper maintenance of oxidoreductase levels is physiologically important. For example, genetically- linked deficiencies in lipoamide dehydrogenase can result in lactic acidosis (Robinson, B.H. et al. (1977) Pediat. Res. 11:1198-1202).
  • Transferases transfer a chemical group from one compound (the donor) to another compound (the acceptor).
  • the types of groups transferred by these enzymes include acyl groups, amino groups, phosphate groups (phosphotransferases or phosphomutases), and others.
  • the transferase carnitine palmitoyltransferase is an important component of fatty acid metabolism. Genetically-Unked deficiencies in this
  • topoisomerases are enzymes that affect the topological state of DNA. For example, defects in topoisomerases or their regulation can affect normal physiology. Reduced levels of topoisomerase II have been correlated with some of
  • Ligases catalyze the formation of a bond between two substrate molecules. The process involves the hydrolysis of a pyrophosphate bond in ATP or a similar energy donor. Ligases are
  • Ligases forming carbon-oxygen bonds include the aminoacyl-transfer RNA (tRNA) synthetases which are important RNA-associated enzymes with roles in translation. Protein biosynthesis depends on each amino acid forming a Unkage with the appropriate tRNA. The 5 aminoacyl-tRNA synthetases are responsible for the activation and correct attachment of an amino acid with its cognate tRNA.
  • the 20 aminoacyl-tRNA synthetase enzymes can be divided into two structural classes, and each class is characterized by a distinctive topology of the catalytic domain. Class I enzymes contain a catalytic domain based on the nucleotide-binding Rossman fold.
  • Class II enzymes contain a central catalytic domain, which consists of a seven-stranded antiparallel ⁇ -sheet 0 motif, as well as N- and C- terminal regulatory domains. Class II enzymes are separated into two groups based on the heterodimeric or homodimeric structure of the enzyme; the latter group is further subdivided by the structure of the N- and C-terminal regulatory domains (Hartlein, M. and S. Cusack (1995) J. Mol. Evol. 40:519-530). Autoantibodies against aminoacyl-tRNAs are generated by patients with dermatomyositis and polymyositis, and correlate strongly with complicating interstitial 5 lung disease (ILD). These antibodies appear to be generated in response to viral infection, and coxsackie virus has been used to induce experimental viral myositis in animals.
  • ILD interstitial 5 lung disease
  • Ligases forming carbon-sulfur bonds mediate a large number of cellular biosynthetic intermediary metabolism processes involve intermolecular transfer of carbon atom-containing substrates (carbon substrates). Examples of such reactions include the tricarboxylic o acid cycle, synthesis of fatty acids and long-chain phospholipids, synthesis of alcohols and aldehydes, synthesis of intermediary metabolites, and reactions involved in the amino acid degradation pathways. Some of these reactions require input of energy, usually in the form of conversion of ATP to either ADP or AMP and pyrophosphate.
  • a carbon substrate is derived from a small molecule containing at least two 5 carbon atoms.
  • the carbon substrate is often covalently bound to a larger molecule which acts as a carbon substrate carrier molecule within the cell.
  • the carrier molecule is coenzyme A.
  • Coenzyme A (CoA) is structurally related to derivatives of the nucleotide ADP and consists of 4'-phosphopantetheine linked via a phosphodiester bond to the alpha phosphate group of adenosine 3',5'-bisphosphate. The terminal thiol group of 4'-phosphopantetheine o acts as the site for carbon substrate bond formation.
  • the predominant carbon substrates which utilizes to the nucleotide ADP and consists of 4'-phosphopantetheine linked via a phosphodiester bond to the alpha phosphate group of adenosine 3',5'-bisphosphate.
  • the terminal thiol group of 4'-phosphopantetheine o
  • CoA as a carrier molecule during biosynthesis and intermediary metabolism in the cell are acetyl, succinyl, and propionyl moieties, collectively referred to as acyl groups.
  • Other carbon substrates include enoyl lipid, which acts as a fatty acid oxidation intermediate, and carnitine, which acts as an acetyl-CoA flux regulator/ mitochondrial acyl group transfer protein.
  • Acyl-CoA and acetyl-CoA are 5 synthesized in the cell by acyl-CoA synthetase and acetyl-CoA synthetase, respectively.
  • acyl-CoA synthetase activity i) acetyl-CoA synthetase, which activates acetate and several other low molecular weight carboxylic acids and is found in muscle mitochondria and the cytosol of other tissues; u) medium-chain acyl-CoA synthetase, which activates fatty acids containing between four and eleven carbon atoms 5 (predominantly from dietary sources), and is present only in liver mitochondria; and in) acyl CoA synthetase, which is specific for long chain fatty acids with between six and twenty carbon atoms, and is found in microsomes and the mitochondria.
  • acyl-CoA synthetase activity has been identified from many sources including bacteria, yeast, plants, mouse, and man.
  • the activity of acyl-CoA synthetase may be modulated by phosphorylation of the enzyme by o cAMP-dependent protein kinase.
  • Ligases forming carbon-nitrogen bonds include amide synthases such as glutamine synthetase (glutamate-ammonia ligase) that catalyzes the animation of glutamic acid to glutamine by ammonia using the energy of ATP hydrolysis.
  • glutamine synthetase glutamine synthetase
  • Glutamine is the primary source for the amino group in various amide transfer reactions involved in de novo pyrimidine nucleotide synthesis and in purine and 5 pyrimidine ribonucleotide interconversions.
  • Overexpression of glutamine synthetase has been observed in primary liver cancer (Christa, L. et al. (1994) Gastroent. 106:1312-1320).
  • Acid-amino-acid ligases are represented by the ubiquitin proteases which are associated with the ubiquitin conjugation system (UCS), a major pathway for the degradation of cellular proteins in eukaryotic cells and some bacteria.
  • UCS ubiquitin conjugation system
  • the UCS mediates the elimination of o abnormal proteins and regulates the half -lives of important regulatory proteins that control cellular processes such as gene transcription and cell cycle progression.
  • proteins targeted for degradation are conjugated to a ubiquitin (Ub), a small heat stable protein.
  • Ub is first activated by a ubiquitin-activating enzyme (El), and then transferred to one of several Ub- conjugating enzymes (E2).
  • E2 then links the Ub molecule through its C-terminal glycine to an 5 internal lysine (acceptor lysine) of a target protein.
  • the ubiquitinated protein is then recognized and degraded by proteasome, a large, multisubunit proteolytic enzyme complex, and ubiquitin is released for reutilization by ubiquitin protease.
  • the UCS is implicated in the degradation of mitotic cyclic kinases, oncoproteins, tumor suppressor genes such as p53, viral proteins, cell surface receptors associated with signal transduction, transcriptional regulators, and mutated or damaged proteins o (Ciechanover, A. (1994) Cell 79: 13-21).
  • a murine proto-oncogene, Unp encodes a nuclear ubiquitin protease whose overexpression leads to oncogenic transformation of NIH3T3 cells, and the human homolog of this gene is consistently elevated in small cell tumors and adenocarcinomas of the lung (Gray, D.A. (1995) Oncogene 10:2179-2183).
  • Cyclo-ligases and other carbon-nitrogen ligases comprise various enzymes and enzyme 5 complexes that participate in the de novo pathways to purine and pyrimidine biosynthesis. Because these pathways are critical to the synthesis of nucleotides for replication of both RNA and DNA, many of these enzymes have been the targets of clinical agents for the treatment of cell proliferative disorders such as cancer and infectious diseases.
  • Purine biosynthesis occurs de novo from the amino acids glycine and glutamine, and other 5 small molecules.
  • Three of the key reactions in this process are catalyzed by a trifunctional enzyme composed of glycinamide-ribonucleotide synthetase (GARS), aminoimidazole ribonucleotide synthetase (AIRS), and glycinamide ribonucleotide transformylase (GART).
  • GART glycinamide ribonucleotide transformylase
  • Adenylosuccinate synthetase catalyzes a later step in purine biosynthesis that converts inosinic acid to adenylosuccinate, a key step on the path to ATP synthesis.
  • This enzyme is also similar to another carbon-nitrogen ligase, argininosuccinate synthetase, that catalyzes a similar reaction in the urea cycle (Powell, S.M. et al. (1992) FEBS Lett. 303:4-10).
  • de novo synthesis of the pyrimidine nucleotides uridylate and cytidylate also arises from a common precursor, in this instance the nucleotide orotidylate derived from orotate and phosphoribosyl pyrophosphate (PPRP).
  • PPRP phosphoribosyl pyrophosphate
  • ATCase aspartate transcarbamylase
  • carbamyl phosphate synthetase II carbamyl phosphate synthetase II
  • DHOase dihydroorotase o
  • Ligases forming carbon-carbon bonds include the carboxylases acetyl-CoA carboxylase and pyruvate carboxylase.
  • Acetyl-CoA carboxylase catalyzes the carboxylation of acetyl-CoA from C0 2 o and H 2 0 using the energy of ATP hydrolysis.
  • Acetyl-CoA carboxylase is the rate-limiting step in the biogenesis of long-chain fatty acids.
  • Two isoforms of acetyl-CoA carboxylase, types I and types II, are expressed in human in a tissue-specific manner (Ha, J. et al. (1994) Eur. J. Biochem. 219:297- 306).
  • Pyruvate carboxylase is a nuclear-encoded mitochondrial enzyme that catalyzes the conversion of pyruvate to oxaloacetate, a key intermediate in the citric acid cycle.
  • 5 Ligases forming phosphoric ester bonds include the DNA ligases involved in both DNA replication and repair.
  • DNA ligases seal phosphodiester bonds between two adjacent nucleotides in a DNA chain using the energy from ATP hydrolysis to first activate the free 5 '-phosphate of one nucleotide and then react it with the 3' -OH group of the adjacent nucleotide.
  • This reseating reaction is used in both DNA replication to join small DNA fragments called Okazaki fragments that are transiently formed in the process of replicating new DNA, and in DNA repair.
  • DNA repair is the process by which accidental base changes, such as those produced by oxidative damage, hydrolytic attack, or uncontrolled methylation of DNA, are corrected before replication or transcription of the DNA can occur.
  • Bloom's syndrome is an inherited human disease in which individuals are partially deficient in DNA Ugation and consequently have an increased incidence of cancer (Alberts, B. et al. (1994) The Molecular Biology of the Cell, Garland Publishing Inc., New York NY, p. 247).
  • SEQ ID NO:69, SEQ ID NO:70, and SEQ ID NO:71 encode, for example, molecules associated with growth and development.
  • Human growth and development requires the spatial and temporal regulation of cell differentiation, cell proliferation, and apoptosis. These processes coordinately control reproduction, aging, embryogenesis, morphogenesis, organogenesis, and tissue repair and maintenance.
  • growth and development is governed by the cell's decision to enter into or exit from the cell division cycle and by the cell's commitment to a terminally differentiated state. These decisions are made by the cell in response to extracellular signals and other environmental cues it receives.
  • the following discussion focuses on the molecular mechanisms of cell division, reproduction, cell differentiation and proUferation, apoptosis, and aging.
  • Cell division is the fundamental process by which all living things grow and reproduce. In unicellular organisms such as yeast and bacteria, each cell division doubles the number of organisms, while in multicellular species many rounds of cell division are required to replace cells lost by wear or by programmed cell death, and for cell differentiation to produce a new tissue or organ. Details of the cell division cycle may vary, but the basic process consists of three principle events. The first event, interphase, involves preparations for cell division, repUcation of the DNA, and production of essential proteins. In the second event, mitosis, the nuclear material is divided and separates to opposite sides of the cell. The final event, cytokinesis, is division and fission of the cell cytoplasm. The sequence and timing of cell cycle transitions is under the control of the cell cycle regulation system which controls the process by positive or negative regulatory circuits at various check points.
  • Regulated progression of the cell cycle depends on the integration of growth control pathways with the basic cell cycle machinery.
  • Cell cycle regulators have been identified by selecting for human and yeast cDNAs that block or activate cell cycle arrest signals in the yeast mating pheromone pathway when they are overexpressed.
  • Known regulators include human CPR (cell cycle progression restoration) genes, such as CPR8 and CPR2, and yeast CDC (cell division control) genes, including CDC91, that block the arrest signals.
  • the CPR genes express a variety of proteins including cycUns, tumor suppressor binding proteins, chaperones, transcription factors, translation factors, and RNA-binding proteins (Edwards, M.C et al.(1997) Genetics 147:1063-1076).
  • Cdks cycUn-dependent kinases
  • the Cdks are composed of a kinase subunit, Cdk, and an activating subunit, cycUn, in a complex that is subject to many levels of regulation.
  • Cdk a kinase subunit
  • cycUn an activating subunit
  • CycUns act by binding to and activating cycUn-dependent protein kinases which then phosphorylate and activate selected proteins involved in the mitotic process.
  • the Cdk-cycUn complex is both positively and negatively regulated by phosphorylation, and by targeted degradation involving molecules such as CDC4 and CDC53.
  • Cdks are further regulated by binding to inhibitors and other proteins such as Sucl that modify their specificity or accessibiUty to regulators (Patra, D. and W.G. Dunphy (1996) Genes Dev. 10:1503-1515; and Mathias, N. et al. (1996) Mol. Cell Biol. 16:6634-6643).
  • Reproduction The male and female reproductive systems are complex and involve many aspects of growth and development. The anatomy and physiology of the male and female reproductive systems are reviewed in (Guyton, A.C. (1991) Textbook of Medical Physiology. W.B. Saunders Co., Philadelphia PA, pp. 899-928).
  • the male reproductive system includes the process of spermatogenesis, in which the sperm are formed, and male reproductive functions are regulated by various hormones and their effects on accessory sexual organs, cellular metabolism, growth, and other bodily functions.
  • Spermatogenesis begins at puberty as a result of stimulation by gonadotropic hormones released from the anterior pituitary. Immature sperm (spermatogonia) undergo several mitotic cell divisions before undergoing meiosis and full maturation. The testes secrete several male sex hormones, the most abundant being testosterone, that is essential for growth and division of the immature sperm, and for the masculine characteristics of the male body. Three other male sex hormones, gonadotropin- releasing hormone (GnRH), luteinizing hormone (LH), and follicle-stimulating hormone (FSH) control sexual function.
  • gonadotropin- releasing hormone GnRH
  • LH luteinizing hormone
  • FSH follicle-stimulating hormone
  • the uterus, ovaries, fallopian tubes, vagina, and breasts comprise the female reproductive system.
  • the ovaries and uterus are the source of ova and the location of fetal development, respectively.
  • the fallopian tubes and vagina are accessory organs attached to the top and bottom of the uterus, respectively.
  • Both the uterus and ovaries have additional roles in the development and loss of reproductive capabiUty during a female' s Ufetime.
  • the primary role of the breasts is lactation. 5
  • Multiple endocrine signals from the ovaries, uterus, pituitary, hypothalamus, adrenal glands, and other tissues coordinate reproduction and lactation. These signals vary during the monthly menstruation cycle and during the female's Ufetime. Similarly, the sensitivity of reproductive organs to these endocrine signals varies during the female's lifetime.
  • a combination of positive and negative feedback to the ovaries, pituitary and hypothalamus o glands controls physiologic changes during the monthly ovulation and endometrial cycles.
  • the anterior pituitary secretes two major gonadotropin hormones, folUcle-stimulating hormone (FSH) and luteinizing hormone (LH), regulated by negative feedback of steroids, most notably by ovarian estradiol. If fertiUzation does not occur, estrogen and progesterone levels decrease. This sudden reduction of the ovarian hormones leads to menstruation, the desquamation of the endometrium. 5 Hormones further govern all the steps of pregnancy, parturition, lactation, and menopause.
  • hCG human chorionic gonadotropin
  • hCS human chorionic somatomammotropin
  • the female breast also matures during pregnancy. Large amounts of estrogen secreted by the placenta trigger growth and branching of the breast milk ductal system while lactation is initiated by the secretion of prolactin by the pituitary gland.
  • Parturition involves several hormonal changes that increase uterine contractility toward the end 5 of pregnancy, as follows.
  • the levels of estrogens increase more than those of progesterone.
  • Oxytocin is secreted by the neurohypophysis. Concomitantly, uterine sensitivity to oxytocin increases.
  • the fetus itself secretes oxytocin, cortisol (from adrenal glands), and prostaglandins.
  • Menopause occurs when most of the ovarian follicles have degenerated.
  • the ovary then produces less estradiol, reducing the negative feedback on the pituitary and hypothalamus glands.
  • o Mean levels of circulating FSH and LH increase, even as ovulatory cycles continue. Therefore, the ovary is less responsive to gonadotropins, and there is an increase in the time between menstrual cycles. Consequently, menstrual bleeding ceases and reproductive capabiUty ends.
  • Tissue growth involves complex and ordered patterns of cell proliferation, cell differentiation, and apoptosis.
  • Cell proliferation must be regulated to maintain both the number of cells and their spatial organization. This regulation depends upon the appropriate expression of proteins which control cell cycle progression in response to extracellular signals, such as growth factors and other mitogens, and intracellular cues, such as DNA damage or nutrient starvation. Molecules which directly or indirectly modulate cell cycle progression fall into several categories, including growth factors and their receptors, second messenger and signal transduction proteins, oncogene products, tumor-suppressor proteins, and mitosis-promoting factors.
  • Growth factors were originally described as serum factors required to promote cell proliferation. Most growth factors are large, secreted polypeptides that act on cells in their local environment. Growth factors bind to and activate specific cell surface receptors and initiate intracellular signal transduction cascades. Many growth factor receptors are classified as receptor tyrosine kinases which undergo autophosphorylation upon ligand binding. Autophosphorylation enables the receptor to interact with signal transduction proteins characterized by the presence of SH2 or SH3 domains (Src homology regions 2 or 3).
  • G-proteins such as Ras, Rab, and Rho
  • GAPs GTPase activating proteins
  • GNRPs guanine nucleotide releasing proteins
  • Small G proteins act as molecular switches that activate other downstream events, such as mitogen-activated protein kinase (MAP kinase) cascades.
  • MAP kinases ultimately activate transcription of mitosis- promoting genes.
  • small signaUng peptides and hormones also influence cell proUferation.
  • GPCR G-protein coupled receptor
  • TGF- ⁇ transforming growth factor beta
  • Some growth factors act on some cells to stimulate cell proliferation and on other cells to inhibit it. Growth factors may also stimulate a cell at one concentration and inhibit the same cell at another concentration. Most growth factors also have a multitude of other actions besides the regulation of cell growth and division: they can control the proUferation, survival, differentiation, migration, or function of cells depending on the circumstance.
  • the tumor necrosis factor/nerve growth factor (TNF/NGF) family can activate or inhibit cell death, as well as regulate proUferation and differentiation.
  • the cell response depends on the type of cell, its stage of differentiation and transformation status, which surface receptors are stimulated, and the types of stimuU acting on the cell (Smith, A. et al. (1994) Cell 76:959-962; and Nocentini, G. et al. (1997) Proc. Natl. Acad. Sci. USA 94:6216-6221).
  • ECM extracellular matrix
  • ECM molecules such as laminin or fibronectin
  • Tenascin-C and -R expressed in developing and lesioned neural tissue, provide stimulatory/anti-adhesive or inhibitory properties, respectively, for axonal growth (Faissner, A. (1997) Cell Tissue Res. 290:331-341).
  • Cancers are associated with the activation of oncogenes which are derived from normal cellular genes. These oncogenes encode oncoproteins which convert normal cells into maUgnant cells. Some oncoproteins are mutant isoforms of the normal protein, and other oncoproteins are abnormally expressed with respect to location or amount of expression.
  • oncoprotein causes cancer by altering transcriptional control of cell proUferation.
  • Five classes of oncoproteins are known to affect cell cycle controls. These classes include growth factors, growth factor receptors, intracellular signal transducers, nuclear transcription factors, and cell-cycle control proteins.
  • Viral oncogenes are integrated into the human genome after infection of human cells by certain viruses. Examples of viral oncogenes include v-src, v-abl, and v-fps.
  • oncogenes have been identified and characterized. These include sis, erbA, erbB, her-2, mutated G s , src, abl, ras, crk, jun, fos, myc, and mutated tumor-suppressor genes such as RB, p53, mdm2, Cipl, pi 6, and cyclin D. Transformation of normal genes to oncogenes may also occur by chromosomal translocation.
  • the Philadelphia chromosome characteristic of chronic myeloid leukemia and a subset of acute lymphoblastic leukemias, results from a reciprocal translocation between chromosomes 9 and 22 that moves a truncated portion of the proto-oncogene c-abl to the breakpoint cluster region (bcr) on chromosome 22.
  • Tumor-suppressor genes are involved in regulating cell proUferation. Mutations which cause reduced or loss of function in tumor-suppressor genes result in uncontrolled cell proUferation.
  • the retinoblastoma gene product RB
  • RB retinoblastoma gene product
  • Phosphorylation of RB causes it to dissociate from the genes, releasing the suppression, and allowing cell division to proceed. Apoptosis
  • Apoptosis is the genetically controlled process by which unneeded or defective cells undergo programmed cell death. Selective eUmination of cells is as important for morphogenesis and tissue 5 remodeUng as is cell proliferation and differentiation. Lack of apoptosis may result in hyperplasia and other disorders associated with increased cell proUferation. Apoptosis is also a critical component of the immune response. Immune cells such as cytotoxic T-cells and natural killer cells prevent the spread of disease by inducing apoptosis in tumor cells and virus-infected cells. In addition, immune cells that fail to distinguish self molecules from foreign molecules must be eUminated by apoptosis to avoid an o autoimmune response.
  • SEQ ID NO:64, SEQ ID NO:65, SEQ ID NO:66, SEQ ID NO:67, and SEQ ID NO:68 encode, for example, biochemical pathway molecules.
  • Biochemical pathways are responsible for regulating metaboUsm, growth and development, protein secretion and trafficking, environmental responses, and ecological interactions including immune response and response to parasites.
  • DNA Deoxyribonucleic acid
  • the bulk of human DNA is nuclear, in the form of Unear chromosomes, while mitochondrial DNA is circular.
  • DNA repUcation begins at specific sites called origins of repUcation. Bidirectional synthesis occurs from the origin via two growing forks that move in opposite directions. Replication is semi-conservative, with each daughter duplex containing one old strand and its newly synthesized complementary partner. Proteins involved in DNA repUcation include DNA polymerases, DNA primase, telomerase, DNA heUcase, topoisomerases, DNA Ugases, repUcation factors, and DNA-binding proteins.
  • DNA Recombination and Repair Cells are constantly faced with repUcation errors and environmental assault (such as ultraviolet irradiation) that can produce DNA damage.
  • Damage to DNA consists of any change that modifies the structure of the molecule. Changes to DNA can be divided into two general classes, single base changes and structural distortions. Any damage to DNA can produce a mutation, and the mutation may produce a disorder, such as cancer. Changes in DNA are recognized by repair systems within the cell. These repair systems act to correct the damage and thus prevent any deleterious affects of a mutational event. Repair systems can be divided into three general types, direct repair, excision repair, and retrieval systems.
  • Proteins involved in DNA repair include DNA polymerase, excision repair proteins, excision and cross Unk repair proteins, recombination and repair proteins, RAD51 proteins, and BLN and WRN proteins that are homologs of RecQ heUcase.
  • DNA polymerase DNA polymerase
  • excision repair proteins excision and cross Unk repair proteins
  • recombination and repair proteins RAD51 proteins
  • BLN and WRN proteins that are homologs of RecQ heUcase.
  • environmental mutagens such as ultraviolet irradiation.
  • Patients with disorders associated with a loss in DNA repair systems often exhibit a high sensitivity to environmental mutagens. Examples of such disorders include xeroderma pigmentosum (XP), Bloom's syndrome (BS), and Werner's syndrome (WS) (Yamagata, K et al. (1998) Proc. Natl. Acad. Sci. USA 95:8733-8738), ataxia telangiectasia, Cockayne's syndrome, and Fanconi
  • Recombination is the process whereby new DNA sequences are generated by the movements of large pieces of DNA.
  • homologous recombination which occurs during meiosis and DNA repair, parent DNA duplexes aUgn at regions of sequence similarity, and new DNA molecules form by the breakage and joining of homologous segments.
  • Proteins involved include RAD51 recombinase.
  • site-specific recombination two specific but not necessarily homologous DNA sequences are exchanged.
  • this process generates a diverse collection of antibody and T cell receptor genes.
  • Proteins involved in site-specific recombination in the immune system include recombination activating genes 1 and 2 (RAG1 and RAG2).
  • a defect in immune system site-specific recombination causes 5 severe combined immunodeficiency disease in mice.
  • RNA Ribonucleic acid
  • ATP ATP
  • CTP CTP
  • UTP UTP
  • GTP GTP
  • RNA Ribonucleic acid
  • ATP ATP
  • CTP CTP
  • UTP UTP
  • GTP GTP
  • RNA Ribonucleic acid
  • RNA is transcribed as a copy of DNA, the genetic material of the organism.
  • RNA rather than DNA serves as the genetic material.
  • RNA copies of the o genetic material encode proteins or serve various structural, catalytic, or regulatory roles in organisms.
  • RNA is classified according to its cellular locaUzation and function.
  • Messenger RNAs encode polypeptides.
  • Ribosomal RNAs (rRNAs) are assembled, along with ribosomal proteins, into ribosomes, which are cytoplasmic particles that translate mRNA into polypeptides.
  • Transfer RNAs (tRNAs) are cytosoUc adaptor molecules that function in mRNA translation by recognizing both an 5 mRNA codon and the amino acid that matches that codon.
  • hnRNAs include mRNA precursors and other nuclear RNAs of various sizes.
  • RNA Transcription o The transcription process synthesizes an RNA copy of DNA. Proteins involved include multi- subunit RNA polymerases, transcription factors IIA, IIB, IID, HE, IIF, IIH, and ILL Many transcription factors incorporate DNA-binding structural motifs which comprise either ⁇ -helices or ⁇ - sheets that bind to the major groove of DNA. Four well-characterized structural motifs are helix-turn- helix, zinc finger, leucine zipper, and helix-loop-helix. 5 RNA Processing
  • RNA processing steps include capping at the 5' end with methylguanosine, polyadenylating the 3' end, and spUcing to remove introns.
  • the spUceosomal complex is comprised of five small nuclear ribonucleoprotein particles (snRNPs) designated Ul, U2, U4, U5, and U6.
  • snRNPs small nuclear ribonucleoprotein particles
  • Ul small nuclear ribonucleoprotein particles
  • snRNP proteins are found in the blood of patients with systemic lupus erythematosus (Stryer, L. (1995) Biochemistry W.H. Freeman and Company, New York NY, p. 863).
  • Heterogeneous nuclear ribonucleoproteins (hnRNPs) have been identified that have roles in spUcing, exporting of the mature RNAs to the cytoplasm, and mRNA translation (Biamonti, G. et al. (1998) CUn. Exp. Rheumatol. 16:317-326).
  • hnRNPs include the yeast proteins Hrplp, involved in cleavage and polyadenylation at the 3' end of the RNA; Cbp80p, involved in 5 capping the 5 ' end of the RNA; and Npl3p, a homolog of mammaUan hnRNP Al , involved in export of mRNA from the nucleus (Shen, E.C et al. (1998) Genes Dev. 12:679-691). HnRNPs have been shown to be important targets of the autoimmune response in rheumatic diseases (Biamonti, supra).
  • RNA recognition motif (Reviewed in Birney, E. et al. (1993) Nucleic Acids Res. 21 :5803- 0 5816.)
  • the RRM is about 80 amino acids in length and forms four ⁇ -strands and two ⁇ -helices arranged in an ⁇ / ⁇ sandwich.
  • the RRM contains a core RNP-1 octapeptide motif along with surrounding conserved sequences.
  • RNA heUcases alter and regulate RNA conformation and secondary structure by using energy 5 derived from ATP hydrolysis to destabiUze and unwind RNA duplexes.
  • the most well-characterized and ubiquitous family of RNA heUcases is the DEAD-box family, so named for the conserved B-type ATP-binding motif which is diagnostic of proteins in this family.
  • DEAD-box heUcases Over 40 DEAD-box heUcases have been identified in organisms as diverse as bacteria, insects, yeast, amphibians, mammals, and plants. DEAD-box heUcases function in diverse processes such as translation initiation, spUcing, ribosome o assembly, and RNA editing, transport, and stabiUty.
  • Some DEAD-box heUcases play tissue- and stage- specific roles in spermatogenesis and embryogenesis. (Reviewed in Linder, P. et al. (1989) Nature 337:121-122.)
  • DEAD-box 1 protein may play a role in the progression of neuroblastoma (Nb) and retinoblastoma (Rb) tumors.
  • DEAD-box heUcases have been implicated 5 either directly or indirectly in ultraviolet Ught-induced tumors, B cell lymphoma, and myeloid maUgnancies. (Reviewed in Godbout, R. et al. (1998) J. Biol. Chem. 273:21161-21168.)
  • RNases Ribonucleases catalyze the hydrolysis of phosphodiester bonds in RNA chains, thus cleaving the RNA.
  • RNase P is a ribonucleoprotein enzyme which cleaves the 5' end of pre-tRNAs as part of their maturation process.
  • RNase H digests the RNA strand of an RNA/DNA o hybrid. Such hybrids occur in cells invaded by retroviruses, and RNase H is an important enzyme in the retroviral repUcation cycle.
  • RNase H domains are often found as a domain associated with reverse transcriptases.
  • RNase activity in serum and cell extracts is elevated in a variety of cancers and infectious diseases (Schein, CH. (1997) Nat. Biotechnol. 15:529-536). Regulation of RNase activity is being investigated as a means to control tumor angiogenesis, allergic reactions, viral infection and repUcation, and fungal infections. Protein Translation
  • the eukaryotic ribosome is composed of a 60S (large) subunit and a 40S (small) subunit, which together form the 80S ribosome.
  • the ribosome 5 also contains more than fifty proteins.
  • the ribosomal proteins have a prefix which denotes the subunit to which they belong, either L (large) or S (small).
  • L (large) or S (small) Three important sites are identified on the ribosome.
  • the aminoacyl-tRNA site (A site) is where charged tRNAs (with the exception of the initiator-tRNA) bind on arrival at the ribosome.
  • the peptidyl-tRNA site (P site) is where new peptide bonds are formed, as well as where the initiator tRNA binds.
  • the exit site (E site) is where deacylated tRNAs o bind prior to their release from the ribosome.
  • Protein biosynthesis depends on each amino acid forming a linkage with the appropriate tRNA. 5
  • the aminoacyl-tRNA synthetases are responsible for the activation and correct attachment of an amino acid with its cognate tRNA.
  • the 20 aminoacyl-tRNA synthetase enzymes can be divided into two structural classes, Class I and Class II. Autoantibodies against aminoacyl-tRNAs are generated by patients with dermatomyositis and polymyositis, and correlate strongly with compUcating interstitial lung disease (ILD). These antibodies appear to be generated in response to viral infection, and o coxsackie virus has been used to induce experimental viral myositis in animals.
  • ILD interstitial lung disease
  • Initiation of translation can be divided into three stages.
  • the first stage brings an initiator transfer RNA (Met-tRNA f ) together with the 40S ribosomal subunit to form the 43S preinitiation complex.
  • the second stage binds the 43S preinitiation complex to the mRNA, followed by migration of 5 the complex to the correct AUG initiation codon.
  • the third stage brings the 60S ribosomal subunit to the 40S subunit to generate an 80S ribosome at the initiation codon.
  • Regulation of translation primarily involves the first and second stage in the initiation process (Pain, V.M. (1996) Eur. J. Biochem. 236:747-771).
  • eIF2 a guanine nucleotide binding protein, recruits the initiator tRNA to the 40S ribosomal subunit. Only when eIF2 is bound to GTP does it associate with the initiator tRNA.
  • eIF2B a guanine nucleotide exchange protein, is responsible for converting eIF2 from the GDP-bound inactive form to the GTP-bound active form.
  • elFl A and eIF3 bind and stabilize the 40S subunit by interacting with 18S ribosomal RNA and specific ribosomal structural proteins.
  • eIF3 is also involved in association of the 40S ribosomal subunit with mRNA.
  • the Met-tRNA f , elFl A, eIF3, and 40S ribosomal subunit together make up the 43 S preinitiation complex (Pain, supra).
  • eIF4F is a complex consisting of three proteins: eIF4E, eIF4A, and eIF4G.
  • eIF4E recognizes and binds to the mRNA 5 -terminal m 7 GTP cap
  • eIF4A is a bidirectional RNA-dependent heUcase
  • eIF4G is a scaffolding polypeptide.
  • eIF4G has three binding domains.
  • eIF4G acts as a bridge between the 40S ribosomal subunit and the mRNA (Hentze, M.W. (1997) Science 275:500-501).
  • the abiUty of eIF4F to initiate binding of the 43S preinitiation complex is regulated by structural features of the mRNA.
  • the mRNA molecule has an untranslated region (UTR) between the 5' cap and the AUG start codon. In some mRNAs this region forms secondary structures that impede binding of the 43S preinitiation complex.
  • UTR untranslated region
  • eIF4A The heUcase activity of eIF4A is thought to function in removing this secondary structure to faciUtate binding of the 43S preinitiation complex (Pain, supra).
  • Elongation is the process whereby additional amino acids are joined to the initiator methionine to form the complete polypeptide chain.
  • the elongation factors EFl ⁇ , EFl ⁇ ⁇ , and EF2 are involved in elongating the polypeptide chain following initiation.
  • EF 1 ⁇ is a GTP-binding protein. In EF 1 ⁇ ' s GTP-bound form, it brings an aminoacyl-tRNA to the ribosome' s A site. The amino acid attached to the newly arrived aminoacyl-tRNA forms a peptide bond with the initiator methionine.
  • the GTP on EFl ⁇ is hydrolyzed to GDP, and EFl ⁇ -GDP dissociates from the ribosome.
  • EFl ⁇ ⁇ binds EFl ⁇ -GDP and induces the dissociation of GDP from EFl ⁇ , allowing EFl ⁇ to bind GTP and a new cycle to begin.
  • EF-G another GTP-binding protein, catalyzes the translocation of tRNAs from the A site to the P site and finally to the E site of the ribosome. This allows the processivity of translation. Translation Termination
  • the release factor eRF carries out termination of translation. eRF recognizes stop codons in the mRNA, leading to the release of the polypeptide chain from the ribosome.
  • Proteins may be modified after translation by the addition of phosphate, sugar, prenyl, fatty acid, and other chemical groups. These modifications are often required for proper protein activity. Enzymes involved in post-translational modification include kinases, phosphatases, glycosyltransferases, and prenyltransferases. The conformation of proteins may also be modified after translation by the introduction and rearrangement of disulfide bonds (rearrangement catalyzed by protein disulfide isomerase), the isomerization of proline sidechains by prolyl isomerase, and by interactions with molecular chaperone proteins. Proteins may also be cleaved by proteases.
  • proteases include serine proteases, cysteine proteases, aspartic proteases, and metalloproteases.
  • Signal peptidase in the endoplasmic reticulum (ER) lumen cleaves the signal peptide from membrane or secretory proteins that are imported into the ER.
  • Ubiquitin proteases are associated with the ubiquitin conjugation system (UCS), a major pathway for the degradation of cellular proteins in eukaryotic cells and some bacteria. The UCS mediates the elimination of abnormal proteins and regulates the half-Uves of important regulatory proteins that control cellular processes such as gene transcription and cell cycle progression.
  • UCS ubiquitin conjugation system
  • proteins targeted for degradation are conjugated to a ubiquitin, a small heat stable protein.
  • Proteins involved in the UCS include ubiquitin-activating enzyme, ubiquitin-conjugating enzymes, ubiquitin-ligases, and ubiquitin C-terminal hydrolases.
  • the ubiquitinated protein is then recognized and degraded by the proteasome, a large, multisubunit proteolytic enzyme complex, and ubiquitin is released for reutilization by ubiquitin protease.
  • Lipids are water-insoluble, oily or greasy substances that are soluble in nonpolar solvents such as chloroform or ether.
  • Neutral fats triacylglycerols serve as major fuels and energy stores.
  • Upids such as phosphoUpids, sphingoUpids, glycoUpids, and cholesterol, are key structural components of cell membranes.
  • Lipid metabolism is involved in human diseases and disorders.
  • atherosclerosis fatty lesions form on the inside of the arterial wall. These lesions promote the loss of arterial flexibiUty and the formation of blood clots (Guyton, A.C. Textbook of Medical Physiology
  • Niemann-Pick diseases types A and B are caused by accumulation of sphingomyelin (a sphingoUpid) and other Upids in the central nervous system due to a defect in the enzyme sphingomyeUnase, leading to neurodegeneration and lung disease.
  • Niemann-Pick disease type C results from a defect in cholesterol transport, leading to the accumulation of sphingomyelin and cholesterol in lysosomes and a secondary reduction in sphingomyeUnase activity.
  • Neurological symptoms such as grand mal seizures, ataxia, and loss of previously learned speech, manifest 1-2 years after birth.
  • Fatty acids are long-chain organic acids with a single carboxyl group and a long non-polar hydrocarbon tail.
  • Long-chain fatty acids are essential components of glycoUpids, phosphoUpids, and cholesterol, which are building blocks for biological membranes, and of triglycerides, which are biological fuel molecules.
  • Long-chain fatty acids are also substrates for eicosanoid production, and are important in the functional modification of certain complex carbohydrates and proteins. 16-carbon and 18-carbon fatty acids are the most common.
  • Fatty acid synthesis occurs in the cytoplasm. In the first step, acetyl-Coenzyme A (CoA) carboxylase (ACC) synthesizes malonyl-CoA from acetyl-CoA and bicarbonate.
  • CoA acetyl-Coenzyme A
  • ACC carboxylase
  • FAS fatty acid synthase
  • FAS catalyzes the synthesis of palmitate from acetyl-CoA and malonyl-CoA.
  • FAS contains acetyl transferase, malonyl transferase, ⁇ -ketoacetyl synthase, acyl carrier protein, ⁇ -ketoacyl reductase, dehydratase, enoyl reductase, and thioesterase activities.
  • the final product of the FAS reaction is the 16-carbon fatty acid palmitate.
  • Triacylglycerols also known as triglycerides and neutral fats, are major energy stores in animals. Triacylglycerols are esters of glycerol with three fatty acid chains. Glyce ⁇ ol-3-phosphate is produced from dihydroxyacetone phosphate by the enzyme glycerol phosphate dehydrogenase or from glycerol by glycerol kinase. Fatty acid-CoA's are produced from fatty acids by fatty acyl-CoA synthetases. Glyercol-3-phosphate is acylated with two fatty acyl-CoA's by the enzyme glycerol phosphate acyltransferase to give phosphatidate.
  • Phosphatidate phosphatase converts phosphatidate to diacylglycerol, which is subsequently acylated to a triacylglyercol by the enzyme diglyceride acyltransferase.
  • Phosphatidate phosphatase and diglyceride acyltransferase form a triacylglyerol synthetase complex bound to the ER membrane.
  • a major class of phospholipids are the phosphoglycerides, which are composed of a glycerol backbone, two fatty acid chains, and a phosphorylated alcohol.
  • Phosphoglycerides are components of cell membranes.
  • Principal phosphoglycerides are phosphatidyl choUne, phosphatidyl ethanolamine, phosphatidyl serine, phosphatidyl inositol, and diphosphatidyl glycerol. Many enzymes involved in 5 phosphoglyceride synthesis are associated with membranes (Meyers, R.A. (1995) Molecular Biology and Biotechnology, VCH PubUshers Inc., New York NY, pp. 494-501). Phosphatidate is converted to CDP-diacylglycerol by the enzyme phosphatidate cytidylyltransferase (ExPASy ENZYME EC 2.7.7.41).
  • the enzyme phosphatidyl serine decarboxylase catalyzes the conversion of phosphatidyl serine to phosphatidyl ethanolamine, using a pyruvate cofactor (Voelker, D.R. (1997) Biochim. Biophys. Acta 1348:236-244).
  • Phosphatidyl choUne is formed using diet-derived choUne by the reaction of CDP-choUne with 1 ,2- 5 diacylglycerol, catalyzed by diacylglycerol choUnephosphotransferase (ExPASy ENZYME 2.7.8.2).
  • Cholesterol composed of four fused hydrocarbon rings with an alcohol at one end, moderates the fluidity of membranes in which it is inco ⁇ orated.
  • cholesterol is used in the synthesis of steroid hormones such as cortisol, progesterone, estrogen, and testosterone.
  • Bile salts derived from o cholesterol facilitate the digestion of Upids.
  • Cholesterol in the skin forms a barrier that prevents excess water evaporation from the body.
  • Farnesyl and geranylgeranyl groups which are derived from cholesterol biosynthesis intermediates, are post-translationally added to signal transduction proteins such as ras and protein-targeting proteins such as rab. These modifications are important for the activities of these proteins (Guyton, supra; Stryer, supra, pp.
  • HMG-CoA hydroxymethylglutaryl-CoA
  • the rate-Umiting step is the conversion of HMG-CoA to mevalonate by HMG-CoA reductase.
  • the drug lovastatin, a potent inhibitor of HMG-CoA reductase, is given to patients to reduce their serum cholesterol levels.
  • mevalonate pathway enzymes include mevalonate kinase, phosphomevalonate kinase, diphosphomevalonate decarboxylase, isopentenyldiphosphate isomerase, dimethylallyl transferase, geranyl transferase, farnesyl-diphosphate farnesyltransferase, squalene monooxygenase, lanosterol synthase, lathosterol oxidase, and 7-dehydrocholesterol reductase.
  • Cholesterol is used in the synthesis of steroid hormones such as cortisol, progesterone, aldosterone, estrogen, and testosterone.
  • cholesterol is converted to pregnenolone by cholesterol monooxygenases.
  • the other steroid hormones are synthesized from pregnenolone by a series of 5 enzyme-catalyzed reactions including oxidations, isomerizations, hydroxylations, reductions, and demethylations. Examples of these enzymes include steroid ⁇ -isomerase, 3 ⁇ -hydroxy- ⁇ 5 -steroid dehydrogenase, steroid 21 -monooxygenase, steroid 19-hydroxylase, and 3 ⁇ -hydroxysteroid dehydrogenase. Cholesterol is also the precursor to vitamin D.
  • Isoprenoid groups are found in vitamin K, ubiquinone, retinal, doUchol phosphate (a carrier of oUgosaccharides needed for N-linked glycosylation), and farnesyl and geranylgeranyl groups that modify proteins. Enzymes involved include farnesyl transferase, polyprenyl transferases, doUchyl phosphatase, and dolichyl kinase.
  • SphingoUpid MetaboUsm 5 Sphingohpids are an important class of membrane lipids that contain sphingosine, a long chain amino alcohol.
  • sphingolipids are composed of one long-chain fatty acid, one polar head alcohol, and sphingosine or sphingosine derivative.
  • the three classes of sphingolipids are sphingomyeUns, cerebrosides, and gangUosides.
  • SphingomyeUns which contain phosphochoUne or phosphoethanolamine as their head group, are abundant in the myeUn sheath surrounding nerve cells.
  • o Galactocerebrosides which contain a glucose or galactose head group, are characteristic of the brain.
  • Other cerebrosides are found in nonneural tissues.
  • GangUosides whose head groups contain multiple sugar units, are abundant in the brain, but are also found in nonneural tissues.
  • Sphingolipids are built on a sphingosine backbone.
  • Sphingosine is acylated to ceramide by the enzyme sphingosine acetyltransferase.
  • Ceramide and phosphatidyl choUne are converted to 5 sphingomyeUn by the enzyme ceramide choUne phosphotransferase.
  • Cerebrosides are synthesized by the Unkage of glucose or galactose to ceramide by a transferase. Sequential addition of sugar residues to ceramide by transferase enzymes yields gangUosides. Eicosanoid MetaboUsm
  • Eicosanoids including prostaglandins, prostacycUn, thromboxanes, and leukotrienes, are 20- o carbon molecules derived from fatty acids. Eicosanoids are signaling molecules which have roles in pain, fever, and inflammation. The precursor of all eicosanoids is arachidonate, which is generated from phospholipids by phosphoUpase A 2 and from diacylglycerols by diacylglycerol Upase. Leukotrienes are produced from arachidonate by the action of Upoxygenases. Prostaglandin synthase, reductases, and isomerases are responsible for the synthesis of the prostaglandins.
  • Prostaglandins have roles in inflammation, blood flow, ion transport, synaptic transmission, and sleep.
  • ProstacycUn and the thromboxanes are derived from a precursor prostaglandin by the action of prostacycUn synthase and thromboxane synthases, respectively.
  • Ketone Body MetaboUsm Pairs of acetyl-CoA molecules derived from fatty acid oxidation in the liver can condense to form acetoacetyl-CoA, which subsequently forms acetoacetate, D-3-hydroxybutyrate, and acetone. These three products are known as ketone bodies.
  • Enzymes involved in ketone body metabolism include HMG-CoA synthetase, HMG-CoA cleavage enzyme, D-3-hydroxybutyrate dehydrogenase, acetoacetate decarboxylase, and 3-ketoacyl-CoA transferase.
  • Ketone bodies are a normal fuel supply of the heart and renal cortex. Acetoacetate produced by the liver is transported to cells where the acetoacetate is converted back to acetyl-CoA and enters the citric acid cycle. In times of starvation, ketone bodies produced from stored triacylglyerols become an important fuel source, especially for the brain. Abnormally high levels of ketone bodies are observed in diabetics. Diabetic coma can result if ketone body levels become too great. Lipid MobiUzation
  • Diazepam binding inhibitor also known as endozepine and acyl CoA-binding protein, is an endogenous ⁇ -aminobutyric acid (GABA) receptor Ugand which is thought to down-regulate the effects of GABA.
  • DBI binds medium- and long-chain acyl-CoA esters with very high affinity and may function as an intracellular carrier of acyl-CoA esters (OMIM *125950 Diazepam Binding Inhibitor; DBI; PROSITE PDOC00686 Acyl-CoA-binding protein signature).
  • Fat stored in Uver and adipose triglycerides may be released by hydrolysis and transported in the blood. Free fatty acids are transported in the blood by albumin. Triacylglycerols and cholesterol esters in the blood are transported in Upoprotein particles.
  • the particles consist of a core of hydrophobic lipids surrounded by a shell of polar Upids and apoUpoproteins.
  • the protein components serve in the solubilization of hydrophobic Upids and also contain cell-targeting signals.
  • Lipoproteins include chylomicrons, chylomicron remnants, very-low-density Upoproteins (VLDL), intermediate- density Upoproteins (IDL), low-density Upoproteins (LDL), and high-density Upoproteins (HDL).
  • VLDL very-low-density Upoproteins
  • IDL intermediate- density Upoproteins
  • LDL low-density Upoproteins
  • HDL high-density Upoproteins
  • Triacylglycerols in chylomicrons and VLDL are hydrolyzed by Upoprotein Upases that line blood vessels in muscle and other tissues that use fatty acids.
  • Cell surface LDL receptors bind LDL particles which are then internalized by endocytosis. Absence of the LDL receptor, the cause of the disease famiUal hypercholesterolemia, leads to increased plasma cholesterol levels and ultimately to atherosclerosis.
  • Plasma cholesteryl ester transfer protein mediates the transfer of cholesteryl esters from HDL to apoUpoprotein B-containing Upoproteins. Cholesteryl ester transfer protein is important in the reverse cholesterol transport system and may play a role in atherosclerosis (Yamashita, S. et al. (1997) Curr. Opin.
  • Macrophage scavenger receptors which bind and internaUze modified Upoproteins, play a role in lipid transport and may contribute to atherosclerosis (Greaves, D.R. et al. (1998) Curr. Opin. Lipidol. 9:425-432).
  • SREBP sterol regulatory element binding protein
  • OSBP oxysterol- binding protein
  • Mitochondria oxidize short-, medium-, and long-chain fatty acids to produce energy for cells.
  • Mitochondrial beta-oxidation is a major energy source for cardiac and skeletal muscle. In liver, it provides ketone bodies to the peripheral circulation when glucose levels are low as in starvation, endurance exercise, and diabetes (Eaton, S. et al. (1996) Biochem. J. 320:345-357).
  • Peroxisomes oxidize medium-, long-, and very-long-chain fatty acids, dicarboxyUc fatty acids, branched fatty acids, prostaglandins, xenobiotics, and bile acid intermediates.
  • the chief roles of peroxisomal beta-oxidation are to shorten toxic UpophiUc carboxylic acids to faciUtate their excretion and to shorten very-long-chain fatty acids prior to mitochondrial beta-oxidation (Mannaerts, G.P. and P.P. van Veldhoven ( 1993) Biochimie 75:147-158).
  • Enzymes involved in beta-oxidation include acyl CoA synthetase, carnitine acyltransferase, acyl CoA dehydrogenases, enoyl CoA hydratases, L-3-hydroxyacyl CoA dehydrogenase, ⁇ -ketothiolase, 2,4-dienoyl CoA reductase, and isomerase.
  • LPLs LysophosphoUpases 5
  • a particular substrate for LPLs lysophosphatidylchoUne, causes lysis of cell membranes when it is formed or imported into a cell.
  • LPLs are regulated by Upid factors including acylcarnitine, arachidonic acid, and phosphatidic acid.
  • the secretory phosphoUpase A 2 (PLA2) superfamily comprises a number of heterogeneous enzymes whose common feature is to hydrolyze the sn-2 fatty acid acyl ester bond of
  • PLA2 activity generates precursors for the biosynthesis of biologically active Upids, hydroxy fatty acids, and platelet-activating factor.
  • PLA2 hydrolysis of the sn-2 ester bond in phosphoUpids generates free fatty acids, such as arachidonic acid and lysophosphoUpids.
  • Carbohydrates including sugars or saccharides, starch, and cellulose, are aldehyde or ketone compounds with multiple hydroxyl groups. The importance of carbohydrate metaboUsm is demonstrated by the sensitive regulatory system in place for maintenance of blood glucose levels. Two pancreatic hormones, insulin and glucagon, promote increased glucose uptake and storage by cells, and increased glucose release from cells, respectively. Carbohydrates have three important roles in
  • carbohydrates are used as energy stores, fuels, and metabolic intermediates. Carbohydrates are broken down to form energy in glycolysis and are stored as glycogen for later use. Second, the sugars deoxyribose and ribose form part of the structural support of DNA and RNA, respectively. Third, carbohydrate modifications are added to secreted and membrane proteins and Upids as they traverse the secretory pathway. Cell surface carbohydrate-containing macromolecules,
  • glycoproteins including glycoproteins, glycoUpids, and transmembrane proteoglycans, mediate adhesion with other cells and with components of the extracellular matrix.
  • the extracellular matrix is comprised of diverse glycoproteins, glycosaminoglycans (GAGs), and carbohydrate-binding proteins which are secreted from the cell and assembled into an organized meshwork in close association with the cell surface.
  • GAGs glycosaminoglycans
  • carbohydrate-binding proteins which are secreted from the cell and assembled into an organized meshwork in close association with the cell surface.
  • the interaction of the cell with the surrounding matrix profoundly influences cell shape, strength, flexibility, motiUty, and adhesion.
  • Carbohydrate metaboUsm is altered in several disorders including diabetes melUtus, 5 hyperglycemia, hypoglycemia, galactosemia, galactokinase deficiency, and UDP-galactose-4-epimerase deficiency (Fauci, AS. et al. (1998) Harrison's Principles of Internal Medicine. McGraw-Hill, New York NY, pp. 2208-2209).
  • Altered carbohydrate metaboUsm is associated with cancer. Reduced GAG and proteoglycan expression is associated with human lung carcinomas (Nackaerts, K. et al. (1997) Int. J. Cancer 74:335-345).
  • the carbohydrate determinants sialyl Lewis A and sialyl Lewis X are 0 frequently expressed on human cancer cells (Kannagi, R. (1997) Glycoconj. J. 14:577-584).
  • Enzymes of the glycolytic pathway convert the sugar glucose to pyruvate while simultaneously o producing ATP.
  • the pathway also provides building blocks for the synthesis of cellular components such as long-chain fatty acids. After glycolysis, pyrvuate is converted to acetyl-Coenzyme A, which, in aerobic organisms, enters the citric acid cycle.
  • Glycolytic enzymes include hexokinase, phosphoglucose isomerase, phosphofructokinase, aldolase, triose phosphate isomerase, glyceraldehyde 3-phosphate dehydrogenase, phosphoglycerate kinase, phosphoglyceromutase, enolase, and pyruvate kinase.
  • phosphofructokinase, hexokinase, and pyruvate kinase are important in regulating the rate of glycolysis.
  • Gluconeogenesis is the synthesis of glucose from noncarbohydrate precursors such as lactate and amino acids.
  • the pathway which functions mainly in times of starvation and intense exercise, o occurs mostly in the liver and kidney.
  • responsible enzymes include pyruvate carboxylase, phosphoenolpyruvate carboxykinase, fructose 1,6-bisphosphatase, and glucose-6-phosphatase. Pentose Phosphate Pathway
  • Pentose phosphate pathway enzymes are responsible for generating the reducing agent NADPH, while at the same time oxidizing glucose-6-phosphate to ribose-5 -phosphate. Ribose-5- phosphate and its derivatives become part of important biological molecules such as ATP, Coenzyme A, NAD + , FAD, RNA, and DNA.
  • the pentose phosphate pathway has both oxidative and non- oxidative branches. The oxidative branch steps, which are catalyzed by the enzymes glucose-6- phosphate dehydrogenase, lactonase, and 6-phosphogluconate dehydrogenase, convert glucose-6- 5 phosphate and NADP + to ribulose-6-phosphate and NADPH.
  • non-oxidative branch steps which are catalyzed by the enzymes phosphopentose isomerase, phosphopentose epimerase, transketolase, and transaldolase, allow the interconversion of three-, four-, five-, six-, and seven-carbon sugars.
  • Glucouronate MetaboUsm isomerase, phosphopentose epimerase, transketolase, and transaldolase
  • Glucuronate is a monosaccharide which, in the form of D-glucuronic acid, is found in the o GAGs chondroitin and dermatan. D-glucuronic acid is also important in the detoxification and excretion of foreign organic compounds such as phenol. Enzymes involved in glucuronate metaboUsm include UDP-glucose dehydrogenase and glucuronate reductase. Disaccharide MetaboUsm
  • Disaccharides must be hydrolyzed to monosaccharides to be digested. Lactose, a disaccharide 5 found in milk, is hydrolyzed to galactose and glucose by the enzyme lactase. Maltose is derived from plant starch and is hydrolyzed to glucose by the enzyme maltase. Sucrose is derived from plants and is hydrolyzed to glucose and fructose by the enzyme sucrase. Trehalose, a disaccharide found mainly in insects and mushrooms, is hydrolyzed to glucose by the enzyme trehalase (OMIM *275360 Trehalase; Ruf, J. et al. (1990) J. Biol. Chem. 265:15034-15039).
  • Lactase, maltase, sucrase, and trehalase are o bound to mucosal cells lining the small intestine, where they participate in the digestion of dietary disaccharides.
  • the enzyme lactose synthetase composed of the catalytic subunit galactosyltransferase and the modifier subunit ⁇ -lactalbumin, converts UDP-galactose and glucose to lactose in the mammary glands.
  • Glycogen is the storage form of carbohydrates in mammals. MobiUzation of glycogen maintains glucose levels between meals and during muscular activity. Glycogen is stored mainly in the Uver and in skeletal muscle in the form of cytoplasmic granules. These granules contain enzymes that catalyze the synthesis and degradation of glycogen, as well as enzymes that regulate these processes. Enzymes that catalyze the degradation of glycogen include glycogen phosphorylase, a transferase, ⁇ - 0 1 ,6-glucosidase, and phosphoglucomutase.
  • Enzymes that catalyze the synthesis of glycogen include UDP-glucose pyrophosphorylase, glycogen synthetase, a branching enzyme, and nucleoside diphosphokinase.
  • the enzymes of glycogen synthesis and degradation are tightly regulated by the hormones insuUn, glucagon, and epinephrine.
  • Starch a plant-derived polysaccharide, is hydrolyzed to maltose, maltotriose, and ⁇ -dextrin by ⁇ -amylase, an enzyme secreted by the saUvary glands and pancreas.
  • Chitin is a polysaccharide found in insects and Crustacea.
  • GAGs Glycosaminoglycans
  • GAGs are anionic Unear unbranched polysaccharides composed of repetitive disaccharide units. These repetitive units contain a derivative of an amino sugar, either glucosamine or galactosamine. GAGs exist free or as part of proteoglycans, large molecules composed of a core protein attached to one or more GAGs.
  • GAGs are found on the cell surface, inside cells, and in the extracellular matrix. Changes in GAG levels are associated with several autoimmune diseases o including autoimmune thyroid disease, autoimmune diabetes melUtus, and systemic lupus erythematosus (Hansen, C et al. (1996) CUn. Exp. Rheum. 14 (Suppl. 15):S59-S67). GAGs include chondroitin sulfate, keratan sulfate, heparin, heparan sulfate, dermatan sulfate, and hyaluronan.
  • HA GAG hyaluronan
  • GAG hyaluronan The GAG hyaluronan (HA) is found in the extracellular matrix of many cells, especially in soft connective tissues, and is abundant in synovial fluid (PitsilUdes, AA. et al. (1993) Int. J. Exp. Pathol. 5 74:27-34). HA seems to play important roles in cell regulation, development, and differentiation (Laurent, T.C. and J.R. Fraser (1992) FASEB J. 6:2397-2404).
  • Hyaluronidase is an enzyme that degrades HA to oUgosaccharides. Hyaluronidases may function in cell adhesion, infection, angiogenesis, signal transduction, reproduction, cancer, and inflammation.
  • Proteoglycans also known as peptidoglycans, are found in the extracellular matrix of o connective tissues such as cartilage and are essential for distributing the load in weight-bearing joints.
  • Cell-surface-attached proteoglycans anchor cells to the extracellular matrix. Both extracellular and cell-surface proteoglycans bind growth factors, facilitating their binding to cell-surface receptors and subsequent triggering of signal transduction pathways.
  • Amino Acid and Nitrogen Metabolism 5 NH 4 + is assimilated into amino acids by the actions of two enzymes, glutamate dehydrogenase and glutamine synthetase. The carbon skeletons of amino acids come from the intermediates of glycolysis, the pentose phosphate pathway, or the citric acid cycle. Of the twenty amino acids used in proteins, humans can synthesize only thirteen (nonessential amino acids). The remaining nine must come from the diet (essential amino acids).
  • Enzymes involved in nonessential o amino acid biosynthesis include glutamate kinase dehydrogenase, pyrroline carboxylate reductase, asparagine synthetase, phenylalanine oxygenase, methionine adenosyltransferase, adenosylhomocysteinase, cystathionine ⁇ -synthase, cystathionine ⁇ -lyase, phosphoglycerate dehydrogenase, phosphoserine transaminase, phosphoserine phosphatase, serine hydroxylmethyltransferase, and glycine synthase.
  • Metabolism of amino acids takes place almost entirely in the liver, where the amino group is removed by aminotransferases (transaminases), for example, alanine aminotransferase.
  • the amino group is transferred to ⁇ -ketoglutarate to form glutamate.
  • Glutamate dehydrogenase converts glutamate to NH 4 + and ⁇ -ketoglutarate.
  • NH 4 + is converted to urea by the urea cycle which is 5 catalyzed by the enzymes arginase, ornithine transcarbamoylase, arginosuccinate synthetase, and arginosuccinase.
  • Carbamoyl phosphate synthetase is also involved in urea formation.
  • Enzymes involved in the metabolism of the carbon skeleton of amino acids include serine dehydratase, asparaginase, glutaminase, propionyl CoA carboxylase, methylmalonyl CoA mutase, branched-chain ⁇ -keto dehydrogenase complex, isovaleryl CoA dehydrogenase, ⁇ -methylcrotonyl CoA carboxylase, o phenylalanine hydroxylase, p-hydroxylphenylpyruvate hydroxylase, and homogentisate oxidase.
  • Polyamines which include spermidine, putrescine, and spermine, bind tightly to nucleic acids and are abundant in rapidly proliferating cells. Enzymes involved in polyamine synthesis include ornithine decarboxylase.
  • MetaboUsm proceeds along separate reaction pathways connected by key intermediates such as acetyl coenzyme A (acetyl-CoA). MetaboUc pathways feature anaerobic and aerobic degradation, coupled with the energy-requiring reactions such as phosphorylation of adenosine diphosphate (ADP) to the triphosphate (ATP) or analogous phosphorylations of guanosine (GDP/GTP), uridine (UDP/UTP), or cytidine (CDP/CTP). Subsequent dephosphorylation of the 5 triphosphate drives reactions needed for cell maintenance, growth, and proliferation.
  • ADP adenosine diphosphate
  • ATP triphosphate
  • UDP/UTP uridine
  • CDP/CTP cytidine
  • Digestive enzymes convert carbohydrates and sugars to glucose; fructose and galactose are converted in the liver to glucose. Enzymes involved in these conversions include galactose- 1- phosphate uridyl transferase and UDP-galactose-4 epimerase.
  • glycolysis converts glucose to pyruvate in a series of reactions coupled to ATP synthesis.
  • o Pyruvate is transported into the mitochondria and converted to acetyl-CoA for oxidation via the citric acid cycle, involving pyruvate dehydrogenase components, dihydrolipoyl transacetylase, and dihydrolipoyl dehydrogenase.
  • Enzymes involved in the citric acid cycle include: citrate synthetase, aconitases, isocitrate dehydrogenase, alpha-ketoglutarate dehydrogenase complex including transsuccinylases, succinyl CoA synthetase, succinate dehydrogenase, fumarases, and malate 5 dehydrogenase.
  • Acetyl CoA is oxidized to C0 2 with concomitant formation of NADH, FADH 2 , and GTP.
  • Enzyme complexes responsible for electron transport and ATP synthesis include the F ⁇ ATPase complex, ubiquinone(CoQ)-cytochrome c reductase, ubiquinone 5 reductase, cytochrome b, cytochrome C j , FeS protein, and cytochrome c oxidase.
  • Triglycerides are hydrolyzed to fatty acids and glycerol by Upases. Glycerol is then phosphorylated to glycerol-3-phosphate by glycerol kinase and glycerol phosphate dehydrogenase, and degraded by the glycolysis. Fatty acids are transported into the mitochondria as fatty acyl- carnitine esters and undergo oxidative degradation. 0 In addition to metaboUc disorders such as diabetes and obesity, disorders of energy metabolism are associated with cancers (Dorward, A. et al. (1997) J. Bioenerg. Biomembr. 29:385- 392), autism (Lombard, J. (1998) Med.
  • Cofactors are small molecular weight inorganic or organic compounds that are required for the action of an enzyme. Many cofactors contain vitamins as a component. Cofactors include thiamine pyrophosphate, flavin adenine dinucleotide, flavin 5 mononucleotide, nicotinamide adenine dinucleotide, pyridoxal phosphate, coenzyme A, tetrahydrofolate, Upoamide, and heme. The vitamins biotin and cobalamin are associated with enzymes as well. Heme, a prosthetic group found in myoglobin and hemoglobin, consists of protopo ⁇ hyrin group bound to iron.
  • Po ⁇ hyrin groups contain four substituted pyrroles covalently joined in a ring, often with a bound metal atom. Enzymes involved in po ⁇ hyrin synthesis include ⁇ - o aminolevulinate synthase, ⁇ -aminolevulinate dehydrase, po ⁇ hobilinogen deaminase, and cosynthase.
  • heme formation causes po ⁇ hyrias. Heme is broken down as a part of erythrocyte turnover. Enzymes involved in heme degradation include heme oxygenase and biliverdin reductase. Iron is a required cofactor for many enzymes. Besides the heme-containing enzymes, iron is found in iron-sulfur clusters in proteins including aconitase, succinate dehydrogenase, and NADH-Q 5 reductase. Iron is transported in the blood by the protein transferrin. Binding of transferrin to the transferrin receptor on cell surfaces allows uptake by receptor mediated endocytosis. Cytosolic iron is bound to ferritin protein.
  • a molybdenum-containing cofactor (molybdopterin) is found in enzymes including sulfite oxidase, xanthine dehydrogenase, and aldehyde oxidase. Molybdopterin biosynthesis is performed by 5 two molybdenum cofactor synthesizing enzymes. Deficiencies in these enzymes cause mental retardation and lens dislocation. Other diseases caused by defects in cofactor metabolism include pernicious anemia and methylmalonic aciduria. Secretion and Trafficking
  • Eukaryotic cells are bound by a Upid bilayer membrane and subdivided into functionally o distinct, membrane bound compartments.
  • the membranes maintain the essential differences between the cytosol, the extracellular environment, and the lumenal space of each intracellular organelle.
  • Upid membranes are highly impermeable to most polar molecules, transport of essential nutrients, metaboUc waste products, cell signaUng molecules, macromolecules and proteins across lipid membranes and between organelles must be mediated by a variety of transport-associated molecules. 5 Protein Trafficking
  • ER-bound ribosomes In eukaryotes, some proteins are synthesized on ER-bound ribosomes, co-translationally imported into the ER, deUvered from the ER to the Golgi complex for post-translational processing and sorting, and transported from the Golgi to specific intracellular and extracellular destinations. All cells possess a constitutive transport process which maintains homeostasis between the cell and its o environment. In many differentiated cell types, the basic machinery is modified to carry out specific transport functions. For example, in endocrine glands, hormones and other secreted proteins are packaged into secretory granules for regulated exocytosis to the cell exterior.
  • the ERGIC matures progressively through the cis, medial, and trans cisternal stacks of the Golgi, modifying the enzyme composition by retrograde transport of specific Golgi enzymes. In this way, proteins moving through the Golgi undergo post-translational modification, such as glycosylation.
  • the final Golgi compartment is the Trans-Golgi Network (TGN), where both membrane and lumenal proteins are sorted for their final destination. Transport vesicles destined for intracellular compartments, such as the lysosome, bud off the TGN.
  • TGN Trans-Golgi Network
  • secretory vesicle which contains proteins destined for the plasma membrane, such as receptors, adhesion molecules, and ion channels, and secretory proteins, such as hormones, neurotransmitters, and digestive 5 enzymes.
  • Secretory vesicles eventually fuse with the plasma membrane (GUck, B.S. and V. Malhotra (1998) Cell 95:883-889).
  • the secretory process can be constitutive or regulated. Most cells have a constitutive pathway for secretion, whereby vesicles derived from maturation of the TGN require no specific signal to fuse with the plasma membrane. In many cells, such as endocrine cells, digestive cells, and neurons, vesicle o pools derived from the TGN collect in the cytoplasm and do not fuse with the plasma membrane until they are directed to by a specific signal.
  • Endocytosis wherein cells internaUze material from the extracellular environment, is essential for transmission of neuronal, metabolic, and proUferative signals; uptake of many essential nutrients; 5 and defense against invading organisms. Most cells exhibit two forms of endocytosis. The first, phagocytosis, is an actin-driven process exempUfied in macrophage and neutrophils. Material to be endocytosed contacts numerous cell surface receptors which stimulate the plasma membrane to extend and surround the particle, enclosing it in a membrane-bound phagosome. In the mammalian immune system, IgG-coated particles bind Fc receptors on the surface of phagocytic leukocytes. Activation of 0 the Fc receptors initiates a signal cascade involving src-family cytosoUc kinases and the monomeric
  • GTP-binding (G) protein Rho The resulting actin reorganization leads to phagocytosis of the particle.
  • This process is an important component of the humoral immune response, allowing the processing and presentation of bacterial-derived peptides to antigen-specific T-lymphocytes.
  • pinocytosis The second form of endocytosis, pinocytosis, is a more generaUzed uptake of material from the 5 external miUeu. Like phagocytosis, pinocytosis is activated by Ugand binding to cell surface receptors.
  • Activation of individual receptors stimulates an internal response that includes coalescence of the receptor-Ugand complexes and formation of clathrin-coated pits.
  • Imagination of the plasma membrane at clathrin-coated pits produces an endocytic vesicle within the cell cytoplasm.
  • These vesicles undergo homotypic fusion to form an early endosomal (EE) compartment.
  • the tubulovesicular EE serves as a o sorting site for incoming material.
  • ATP-driven proton pumps in the EE membrane lowers the pH of the
  • EE lumen (pH 6.3-6.8).
  • the acidic environment causes many Ugands to dissociate from their receptors.
  • the receptors, along with membrane and other integral membrane proteins, are recycled back to the plasma membrane by budding off the tubular extensions of the EE in recycUng vesicles (RV).
  • RV recycUng vesicles
  • This selective removal of recycled components produces a carrier vesicle containing Ugand and other material from the external environment.
  • the carrier vesicle fuses with TGN-derived vesicles which contain hydrolytic enzymes.
  • the acidic environment of the resulting late endosome (LE) activates the hydrolytic enzymes which degrade the Ugands and other material. As digestion takes place, the LE fuses with the lysosome where digestion is completed (MeUman, I. (1996) Annu. Rev. Cell Dev. Biol. 5 12:575-625).
  • RecycUng vesicles may return directly to the plasma membrane.
  • Receptors internaUzed and returned directly to the plasma membrane have a turnover rate of 2-3 minutes.
  • Some RVs undergo microtubule-directed relocation to a perinuclear site, from which they then return to the plasma membrane. Receptors following this route have a turnover rate of 5-10 minutes. Still other RVs are 0 retained within the cell until an appropriate signal is received (Mellman, supra; and James, D.E. et al. (1994) Trends Cell Biol. 4:120-126).
  • vesicles form at the transitional endoplasmic reticulum 5 (tER), the rim of Golgi cisternae, the face of the Trans-Golgi Network (TGN), the plasma membrane (PM), and tubular extensions of the endosomes.
  • tER transitional endoplasmic reticulum 5
  • TGN Trans-Golgi Network
  • PM plasma membrane
  • tubular extensions of the endosomes begins with the budding of a vesicle out of the donor membrane.
  • the membrane-bound vesicle contains proteins to be transported and is surrounded by a protective coat made up of protein subunits recruited from the cytosol.
  • the initial budding and coating processes are controlled by a cytosotic ras-Uke GTP-binding protein, ADP- o ribosylating factor (Arf), and adapter proteins (AP).
  • a cytosotic ras-Uke GTP-binding protein ADP- o ribosylating factor (Arf)
  • AP adapter proteins
  • Different isoforms of both Arf and AP are involved at different sites of budding.
  • Another small G-protein, dynamin forms a ring complex around the neck of the forming vesicle and may provide the mechanochemical force to accompUsh the final step of the budding process.
  • the coated vesicle complex is then transported through the cytosol. During the transport process, Arf-bound GTP is hydrolyzed to GDP and the coat dissociates from the transport 5 vesicle (West, M.A. et al.
  • coat protein Two different classes have also been identified. Clathrin coats form on the TGN and PM surfaces, whereas coatomer or COP coats form on the ER and Golgi. COP coats can further be distinguished as COPI, involved in retrograde traffic through the Golgi and from the Golgi to the ER, and COPII, involved in anterograde traffic from the ER to the Golgi (Mellman, supra).
  • the COP coat consists of two major components, a o G-protein (Arf or Sar) and coat protomer (coatomer).
  • Coatomer is an equimolar complex of seven proteins, termed alpha-, beta-, beta'-, gamma-, delta-, epsilon- and zeta-COP. (Harter, C. and F.T. Wieland (1998) Proc. Natl. Acad. Sci. USA 95:11649-11654.) Membrane Fusion
  • Transport vesicles undergo homotypic or heterotypic fusion in the secretory and endocytotic pathways.
  • Molecules required for appropriate targeting and fusion of vesicles with their target membrane include proteins inco ⁇ orated in the vesicle membrane, the target membrane, and proteins recruited from the cytosol.
  • VAMP vesicle-associated membrane protein
  • a cytosoUc prenylated GTP-binding protein, Rab a member of the Ras superfamily
  • GTP-bound Rab proteins are directed into nascent transport vesicles where they interact with VAMP. Following vesicle transport, GTPase activating proteins (GAPs) in the target membrane convert Rab proteins to the GDP-bound form.
  • GAPs GTPase activating proteins
  • GDI guanine-nucleotide dissociation inhibitor
  • Rab proteins appear to play a role in mediating the function of a viral gene, Rev, which is essential for repUcation of HIV-1, the virus responsible for AIDS (Flavell, R.A. et al. (1996) Proc. Natl. Acad. Sci. USA 93:4421-4424).
  • N-ethylmaleimide sensitive factor (NSF) and soluble NSF-attacbment protein ( ⁇ -SNAP and ⁇ -SNAP) are two such o proteins that are conserved from yeast to man and function in most intracellular membrane fusion reactions.
  • Seel represents a family of yeast proteins that function at many different stages in the secretory pathway including membrane fusion. Recently, mammaUan homologs of Seel, called Munc-18 proteins, have been identified (Katagiri, H. et al. (1995) J. Biol. Chem. 270:4963-4966; Hata et al. supra). 5
  • the SNARE complex involves three SNARE molecules, one in the vesicular membrane and two in the target membrane.
  • Synaptotagmin is an integral membrane protein in the synaptic vesicle which associates with the t-SNARE syntaxin in the docking complex. Synaptotagmin binds calcium in a complex with negatively charged phosphoUpids, which allows the cytosoUc SNAP protein to displace synaptotagmin from syntaxin and fusion to occur. Thus, synaptotagmin is a negative regulator of o fusion in the neuron (Littleton, J.T. et al. (1993) Cell 74:1125-1134). The most abundant membrane protein of synaptic vesicles appears to be the glycoprotein synaptophysin, a 38 kDa protein with four transmembrane domains.
  • v-SNARE v-SNARE
  • t-SNAREs t-SNAREs
  • associated proteins v-SNARE
  • Different isoforms of SNAREs and Rabs show distinct cellular and subcellular distributions.
  • VAMP-1/synaptobrevin, membrane-anchored synaptosome-associated protein of 25 kDa (SNAP- 25), syntaxin-1 , Rab3A, Rabl5, and Rab23 are predominantly expressed in the brain and nervous system.
  • syntaxin, VAMP, and Rab proteins are associated with distinct subcellular compartments and their vesicular carriers. 5 Nuclear Transport
  • NPCs nuclear pore complexes
  • All nuclear proteins are imported from the cytoplasm, their site of synthesis.
  • tRNA and mRNA are exported from the nucleus, their site of synthesis, to the cytoplasm, their site of function.
  • o Processing of small nuclear RNAs involves export into the cytoplasm, assembly with proteins and modifications such as hypermethylation to produce small nuclear ribonuclear proteins (snRNPs), and subsequent import of the snRNPs back into the nucleus.
  • snRNPs small nuclear ribonuclear proteins
  • ribosomes require the initial import of ribosomal proteins from the cytoplasm, their inco ⁇ oration with RNA into ribosomal subunits, and export back to the cytoplasm. (Gorlich, D. and I.W. Mattaj (1996) Science 271:1513- 5 1518.)
  • NLS nuclear locaUzation signals
  • NLS nuclear locaUzation signals
  • NLS are found on proteins that are targeted to the nucleus, such as the glucocorticoid receptor. The NLS is o recognized by the NLS receptor, importin, which then interacts with the monomeric GTP-binding protein Raa
  • This NLS protein/receptor/Ran complex navigates the nuclear pore with the help of the homodimeric protein nuclear transport factor 2 (NTF2).
  • NTF2 binds the GDP-bound form of Ran and to multiple proteins of the nuclear pore complex containing FXFG repeat motifs, such as p62.
  • abnormal hormonal secretion is Unked to disorders such as diabetes insipidus (vasopressin), hyper- and hypoglycemia (insuUn, glucagon), Grave's disease and goiter (thyroid hormone), and Cushing's and Addison's diseases (adrenocorticotropic hormone, ACTH).
  • cancer cells secrete excessive amounts of hormones or other biologically active peptides.
  • Disorders related to excessive secretion of biologically active peptides by tumor cells include fasting hypoglycemia due to increased insuUn secretion from insuUnoma-islet cell tumors; hypertension due to increased epinephrine and norepinephrine secreted from pheochromocytomas of the adrenal medulla and sympathetic paragangUa; and carcinoid syndrome, which is characterized by abdominal cramps, diarrhea, and valvular heart disease caused by excessive amounts of vasoactive substances such as serotonin, bradykinin, histamine, prostaglandins, and polypeptide hormones, secreted from intestinal tumors.
  • vasoactive substances such as serotonin, bradykinin, histamine, prostaglandins, and polypeptide hormones, secreted from intestinal tumors.
  • Biologically active peptides that are ectopically synthesized in and secreted from tumor cells include ACTH and vasopressin (lung and pancreatic cancers); parathyroid hormone (lung and bladder cancers); calcitonin (lung and breast cancers); and thyroid-stimulating hormone (medullary thyroid carcinoma).
  • ACTH and vasopressin lung and pancreatic cancers
  • parathyroid hormone lung and bladder cancers
  • calcitonin lung and breast cancers
  • thyroid-stimulating hormone medullary thyroid carcinoma.
  • Such peptides may be useful as diagnostic markers for tumorigenesis (Schwartz, M.Z. (1997) Semin. Pediatr. Surg. 3:141-146; and Said, S.I. and G.R. Faloona (1975) N. Engl. J. Med. 293:155-160).
  • Defective nuclear transport may play a role in cancer.
  • the BRCAl protein contains three potential NLSs which interact with importin alpha, and is transported into the nucleus by the importin/NPC pathway.
  • the BRCAl protein is aberrantly locahzed in the cytoplasm.
  • the mislocation of the BRCAl protein in breast cancer cells may be due to a defect in the NPC nuclear import pathway (Chen, CF. et al. (1996) J. Biol. Chem. 271:32863-32868).
  • Organisms respond to the environment by a number of pathways.
  • Heat shock proteins including hsp 70, hsp60, hsp90, and hsp 40, assist organisms in coping with heat damage to cellular proteins.
  • Aquaporins are channels that transport water and, in some cases, nonionic small solutes such as urea and glycerol. Water movement is important for a number of physiological processes including renal fluid filtration, aqueous humor generation in the eye, cerebrospinal fluid production in the brain, and appropriate hydration of the lung. Aquaporins are members of the major intrinsic protein (MIP) family of membrane transporters (King, L.S. and P. Agre (1996) Annu. Rev. Physiol. 58:619- 648; Ishibashi, K. et al. (1997) J. Biol. Chem. 272:20782-20786).
  • MIP major intrinsic protein
  • MTs The metallothioneins
  • cysteine-rich proteins that bind heavy metals such as cadmium, zinc, mercury, lead, and copper and are thought to play a role in metal detoxification or the metaboUsm and homeostasis of metals.
  • Arsenite-resistance proteins have been identified in hamsters that are resistant to toxic levels of arsenite (Rossman, T.G. et al. (1997) Mutat. Res. 386:307-314).
  • Proteins involved in light perception include rhodopsin, fransducin, and cGMP phosphodiesterase. Proteins involved in odor perception include multiple olfactory receptors. Other proteins are important in human Circadian rhythms and responses to wounds. Immunity and Host Defense
  • the cellular components of the humoral immune system include six different types of leukocytes: monocytes, lymphocytes, polymo ⁇ honuclear granulocytes (consisting of neutrophils, eosinophils, and basophils) and plasma cells. Additionally, fragments of megakaryocytes, a seventh type of white blood cell in the bone marrow, occur in large numbers in the blood as platelets.
  • Leukocytes are formed from two stem cell lineages in bone marrow.
  • the myeloid stem cell line produces granulocytes and monocytes and, the lymphoid stem cell produces lymphocytes.
  • Lymphoid cells travel to the thymus, spleen and lymph nodes, where they mature and differentiate into lymphocytes.
  • Leukocytes are responsible for defending the body against invading pathogens. Neutrophils and monocytes attack invading bacteria, viruses, and other pathogens and destroy them by phagocytosis.
  • Monocytes enter tissues and differentiate into macrophages which are extremely phagocytic. Lymphocytes and plasma cells are a part of the immune system which recognizes specific foreign molecules and organisms and inactivates them, as well as signals other cells to attack the invaders.
  • Granulocytes and monocytes are formed and stored in the bone marrow until needed. Megakaryocytes are produced in bone marrow, where they fragment into platelets and are released into the bloodstream. The main function of platelets is to activate the blood clotting mechanism. Lymphocytes and plasma cells are produced in various lymphogenous organs, including the lymph nodes, spleen, thymus, and tonsils.
  • Basophils participate in the release of the chemicals involved in the inflammatory process.
  • the main function of basophils is secretion of these chemicals to such a degree that they have been referred to as "unicellular endocrine glands".
  • a distinct aspect of basophilic secretion is that the contents of granules go directly into the extracellular environment, not into vacuoles as occurs with 0 neutrophils, eosinophils and monocytes.
  • Basophils have receptors for the Fc fragment of immunoglobulin E (IgE) that are not present on other leukocytes. CrossUnking of membrane IgE with anti-IgE or other ligands triggers degranulation.
  • IgE immunoglobulin E
  • Eosinophils are bi- or multi-nucleated white blood cells which contain eosinophiUc granules. Their plasma membrane is characterized by Ig receptors, particularly IgG and IgE. Generally, 5 eosinophils are stored in the bone marrow until recruited for use at a site of inflammation or invasion. They have specific functions in parasitic infections and allergic reactions, and are thought to detoxify some of the substances released by mast cells and basophils which cause inflammation. Additionally, they phagocytize antigen-antibody complexes and further help prevent spread of the inflammation.
  • Macrophages are monocytes that have left the blood stream to settle in tissue. Once o monocytes have migrated into tissues, they do not re-enter the bloodstream.
  • the mononuclear phagocyte system is comprised of precursor cells in the bone marrow, monocytes in circulation, and macrophages in tissues. The system is capable of very fast and extensive phagocytosis. A macrophage may phagocytize over 100 bacteria, digest them and extrude residues, and then survive for many more months. Macrophages are also capable of ingesting large particles, including red 5 blood cells and malarial parasites. They increase several-fold in size and transform into macrophages that are characteristic of the tissue they have entered, surviving in tissues for several months.
  • Mononuclear phagocytes are essential in defending the body against invasion by foreign pathogens, particularly intracellular microorganisms such as M. tuberculosis, listeria, leishmania and toxoplasma. Macrophages can also control the growth of tumorous cells, via both phagocytosis and o secretion of hydrolytic enzymes. Another important function of macrophages is that of processing antigen and presenting them in a biochemically modified form to lymphocytes.
  • the immune system responds to invading microorganisms in two major ways: antibody production and cell mediated responses.
  • Antibodies are immunoglobulin proteins produced by B-lymphocytes which bind to specific antigens and cause inactivation or promote destruction of the 5 antigen by other cells.
  • Cell -mediated immune responses involve T-lymphocytes (T cells) that react with foreign antigen on the surface of infected host cells. Depending on the type of T cell, the infected cell is either killed or signals are secreted which activate macrophages and other cells to destroy the infected cell (Paul, supra).
  • T-lymphocytes originate in the bone marrow or liver in fetuses. Precursor cells migrate via 5 the blood to the thymus, where they are processed to mature into T-lymphocytes. This processing is crucial because of positive and negative selection of T cells that will react with foreign antigen and not with self molecules. After processing, T cells continuously circulate in the blood and secondary lymphoid tissues, such as lymph nodes, spleen, certain epithelium-associated tissues in the gastrointestinal tract, respiratory tract and skin. When T-lymphocytes are presented with the o complementary antigen, they are stimulated to proliferate and release large numbers of activated T cells into the lymph system and the blood system. These activated T cells can survive and circulate for several days.
  • T memory cells are created, which remain in the lymphoid tissue for months or years. Upon subsequent exposure to that specific antigen, these memory cells will respond more rapidly and with a stronger response than induced by the original antigen. This creates 5 an "immunological memory” that can provide immunity for years.
  • T cells There are two major types of T cells: cytotoxic T cells destroy infected host cells, and helper T cells activate other white blood cells via chemical signals.
  • helper T cells activates macrophages to destroy ingested microorganisms, while another, T H 2, stimulates the production of antibodies by B cells.
  • T H 1 activates macrophages to destroy ingested microorganisms
  • T H 2 stimulates the production of antibodies by B cells.
  • Cytotoxic T cells directly attack the infected target cell.
  • virus-infected cells peptides derived from viral proteins are generated by the proteasome. These peptides are transported into the ER by the transporter associated with antigen processing (TAP) (Pamer, E. and P. Cresswell (1998) Annu. Rev. Immunol. 16:323-358).
  • TEP antigen processing
  • the peptides bind MHC I chains, and the peptide/MHC I complex is transported to the cell surface.
  • Receptors on the surface of T cells bind to 5 antigen presented on cell surface MHC molecules.
  • T cells secrete ⁇ -interferon, a signal molecule that induces the expression of genes necessary for presenting viral (or other) antigens to cytotoxic T cells. Cytotoxic T cells kill the infected cell by stimulating programmed cell death.
  • Helper T cells constitute up to 75% of the total T cell population. They regulate the immune o functions by producing a variety of lymphokines that act on other cells in the immune system and on bone marrow. Among these lymphokines are: interleukins-2,3,4,5,6; granulocyte-monocyte colony stimulating factor, and ⁇ -interferon.
  • Helper T cells are required for most B cells to respond to antigen.
  • an activated helper cell contacts a B cell, its centrosome and Golgi apparatus become oriented toward the B cell, aiding 5 the directing of signal molecules, such as transmembrane-bound protein called CD40 ligand, onto the B cell surface to interact with the CD40 transmembrane protein.
  • Secreted signals also help B cells to proliferate and mature and, in some cases, to switch the class of antibody being produced.
  • B-lymphocytes produce antibodies which react with specific antigenic proteins presented by pathogens. Once activated, B cells become filled with extensive rough endoplasmic 5 reticulum and are known as plasma cells. As with T cells, interaction of B cells with antigen stimulates proliferation of only those B cells which produce antibody specific to that antigen.
  • Antibodies or immunoglobulins (Ig), are the founding members of the Ig superfamily and the central components of the humoral immune response. Antibodies are either expressed on the surface of B cells or secreted by B cells into the circulation. Antibodies bind and neutralize blood-borne foreign antigens.
  • the prototypical antibody is a tetramer consisting of two identical heavy 5 polypeptide chains (H-chains) and two identical light polypeptide chains (L-chains) interlinked by disulfide bonds. This arrangement confers the characteristic Y-shape to antibody molecules. Antibodies are classified based on their H-chain composition.
  • the five antibody classes, IgA, IgD, IgE, IgG and IgM, are defined by the a, ⁇ , e, ⁇ , and ⁇ H-chain types.
  • IgG the most o common class of antibody found in the circulation, is tetrameric, while the other classes of antibodies are generally variants or multimers of this basic structure.
  • H-chains and L-chains each contain an N-terminal variable region and a C-terminal constant region. Both H-chains and L-chains contain repeated Ig domains. For example, a typical H-chain contains four Ig domains, three of which occur within the constant region and one of which occurs 5 within the variable region and contributes to the formation of the antigen recognition site. Likewise, a typical L-chain contains two Ig domains, one of which occurs within the constant region and one of which occurs within the variable region. In addition, H chains such as ⁇ have been shown to associate with other polypeptides during differentiation of the B cell.
  • Antibodies can be described in terms of their two main functional domains. Antigen o recognition is mediated by the Fab (antigen binding fragment) region of the antibody, while effector functions are mediated by the Fc (crystallizable fragment) region. Binding of antibody to an antigen, such as a bacterium, triggers the destruction of the antigen by phagocytic white blood cells such as macrophages and neutrophils. These cells express surface receptors that specifically bind to the antibody Fc region and allow the phagocytic cells to engulf, ingest, and degrade the antibody-bound 5 antigen.
  • an antigen such as a bacterium
  • the Fc receptors expressed by phagocytic cells are single-pass transmembrane glycoproteins of about 300 to 400 amino acids (Sears, D.W. et al. (1990) J. Immunol. 144:371-378).
  • the extracellular portion of the Fc receptor typically contains two or three Ig domains.
  • AIDS Abnormal Immunodeficiency Syndrome
  • helper T cells are depleted, leaving the patient susceptible to infection by microorganisms and parasites.
  • Another widespread medical condition attributable to the immune system is that of allergic reactions to certain antigens. Allergic reactions include: hay fever, asthma, anaphylaxis, and urticaria (hives).
  • Leukemias are an excess production of white blood cells, to the point where a major portion of the body's metaboUc resources are directed solely at proUferation of white blood cells, leaving other tissues to starve.
  • Leukopenia or agranulocytosis occurs when the bone marrow stops producing white blood cells. This leaves the body unprotected against foreign microorganisms, including those which normally inhabit skin, mucous membranes, and gastrointestinal tract. If all white blood cell production stops completely, infection will occur within two days and death may follow only 1 to 4 days later. Impaired phagocytosis occurs in several diseases, including monocytic leukemia, systemic lupus, and granulomatous disease. In such a situation, macrophages can phagocytize normally, but the enveloped organism is not killed.
  • Eosinophilia is an excess of eosinophils commonly observed in patients with allergies (hay fever, asthma), allergic reactions to drugs, rheumatoid arthritis, and cancers (Hodgkin's disease, lung, and liver cancer) (Isselbacher, KJ. et al. (1994) Harrison's Principles of Internal Medicine, McGraw-Hill, Inc., New York NY).
  • the complement system serves as an effector system and is involved in infectious agent recognition. It can function as an independent immune network or in conjunction with other humoral immune responses.
  • the complement system is comprised of numerous plasma and membrane proteins that act in a cascade of reaction sequences whereby one component activates the next. The result is a rapid and amplified response to infection through either an inflammatory response or increased phagocytosis.
  • the complement system has more than 30 protein components which can be divided into functional groupings including modified serine proteases, membrane-binding proteins and regulators of complement activation. Activation occurs through two different pathways the classical and the alternative. Both pathways serve to destroy infectious agents through distinct triggering mechanisms that eventually merge with the involvement of the component C3.
  • the classical pathway requires antibody binding to infectious agent antigens.
  • the antibodies serve to define the target and initiate the complement system cascade, culminating in the destruction of the infectious agent.
  • the complement can be seen as an effector arm of the humoral immune system.
  • the alternative pathway of the complement system does not require the presence of preexisting antibodies for targeting infectious agent destruction. Rather, this pathway, through low levels of an activated component, remains constantly primed and provides surveillance in the non- immune host to enable targeting and destruction of infectious agents. In this case foreign material triggers the cascade, thereby facilitating phagocytosis or lysis (Paul, supra, pp.918-919).
  • Inflammatory responses are divided into four categories on the basis of pathology and include allergic inflammation, cytotoxic antibody mediated inflammation, immune complex mediated inflammation and monocyte mediated inflammation. Inflammation manifests as a combination of each of these forms with one predominating.
  • Acute inflammation is observed in individuals wherein specific antigens stimulate IgE antibody production.
  • Mast cells and basophils are subsequently activated by the attachment of antigen- IgE complexes, resulting in the release of cytoplasmic granule contents such as histamine.
  • the products of activated mast cells can increase vascular permeability and constrict the smooth muscle of breathing passages, resulting in anaphylaxis or asthma.
  • Acute inflammation is also mediated by cytotoxic antibodies and can result in the destruction of tissue through the binding of complement-fixing antibodies to cells.
  • the responsible antibodies are of the IgG or IgM types. Resultant clinical disorders include autoimmune hemolytic anemia and thrombocytopenia as associated with systemic lupus erythematosis.
  • Immune complex mediated acute inflammation involves the IgG or IgM antibody types which combine with antigen to activate the complement cascade.
  • immune complexes bind to neutrophils and macrophages they activate the respiratory burst to form protein- and vessel- damaging agents such as hydrogen peroxide, hydroxyl radical, hypochlorous acid, and chloramines.
  • Clinical manifestations include rheumatoid arthritis and systemic lupus erythematosus.
  • SEQ ID NO:9 encodes, for example, an extracellular information transmission molecule.
  • Intercellular communication is essential for the growth and survival of multicellular organisms, and in particular, for the function of the endocrine, nervous, and immune systems.
  • intercellular communication is critical for developmental processes such as tissue construction and organogenesis, in which cell proliferation, cell differentiation, and mo ⁇ hogenesis must be spatially and temporally regulated in a precise and coordinated manner.
  • Cells communicate 5 with one another through the secretion and uptake of diverse types of signaling molecules such as hormones, growth factors, neuropeptides, and cytokines. Hormones
  • Hormones are signaUng molecules that coordinately regulate basic physiological processes from embryogenesis throughout adulthood. These processes include metaboUsm, respiration, o reproduction, excretion, fetal tissue differentiation and organogenesis, growth and development, homeostasis, and the stress response. Hormonal secretions and the nervous system are tightly integrated and interdependent. Hormones are secreted by endocrine glands, primarily the hypothalamus and pituitary, the thyroid and parathyroid, the pancreas, the adrenal glands, and the ovaries and testes. The secretion of hormones into the circulation is tightly controlled. Hormones are often 5 secreted in diurnal, pulsatile, and cyclic patterns.
  • Hormone secretion is regulated by perturbations in blood biochemistry, by other upstream-acting hormones, by neural impulses, and by negative feedback loops. Blood hormone concentrations are constantly monitored and adjusted to maintain optimal, steady-state levels. Once secreted, hormones act only on those target cells that express specific receptors. o Most disorders of the endocrine system are caused by either hyposecretion or hypersecretion of hormones. Hyposecretion often occurs when a hormone's gland of origin is damaged or otherwise impaired. Hypersecretion often results from the proliferation of tumors derived from hormone-secreting cells. Inappropriate hormone levels may also be caused by defects in regulatory feedback loops or in the processing of hormone precursors. Endocrine malfunction may also occur when the target cell fails 5 to respond to the hormone.
  • Hormones can be classified biochemically as polypeptides, steroids, eicosanoids, or amines.
  • Polypeptides which include diverse hormones such as insuUn and growth hormone, vary in size and function and are often synthesized as inactive precursors that are processed intracellularly into mature, active forms.
  • Amines which include epinephrine and dopamine, are amino acid derivatives that o function in neuroendocrine signaUng.
  • Steroids which include the cholesterol-derived hormones estrogen and testosterone, function in sexual development and reproduction.
  • Eicosanoids which include prostaglandins and prostacycUns, are fatty acid derivatives that function in a variety of processes.
  • polypeptides and some amines are soluble in the circulation where they are highly susceptible to proteolytic degradation within seconds after their secretion. Steroids and Upids are insoluble and must be transported in the circulation by carrier proteins. The following discussion will focus primarily on polypeptide hormones.
  • Hypothalamic hormones include thyrotropin-releasing hormone, gonadotropin-releasing hormone, somatostatin, growth-hormone releasing factor, corticotropin-releasing hormone, substance P, dopamine, and prolactin-releasing hormone. These hormones directly regulate the secretion of hormones from the anterior lobe of the pituitary.
  • Hormones secreted by the anterior pituitary include adrenocorticotropic hormone (ACTH), melanocyte-stimulating hormone, somatotropic hormones such as growth hormone and prolactin, glycoprotein hormones such as thyroid-stimulating hormone, luteinizing hormone (LH), and folUcle-stimulating hormone (FSH), ⁇ -Upotropin, and ⁇ -endo ⁇ hins.
  • ACTH adrenocorticotropic hormone
  • melanocyte-stimulating hormone such as growth hormone and prolactin
  • glycoprotein hormones such as thyroid-stimulating hormone, luteinizing hormone (LH), and folUcle-stimulating hormone (FSH), ⁇ -Upotropin, and ⁇ -endo ⁇ hins.
  • FSH folUcle-stimulating hormone
  • ⁇ -Upotropin ⁇ -Upotropin
  • disorders of the hypothalamus and pituitary often result from lesions such as primary brain tumors, adenomas, infarction associated with pregnancy, hypophysectomy, aneurysms, vascular malformations, thrombosis, infections, immunological disorders, and compUcations due to head trauma. Such disorders have profound effects on the function of other endocrine glands.
  • Disorders associated with hypopituitarism include hypogonadism, Sheehan syndrome, diabetes insipidus, Kallman's disease, Hand-Schuller-Christian disease, Letterer-Siwe disease, sarcoidosis, empty sella syndrome, and dwarfism.
  • Disorders associated with hype ⁇ ituitarism include acromegaly, giantism, and syndrome of inappropriate ADH secretion (SIADH), often caused by benign adenomas.
  • SIADH inappropriate ADH secretion
  • Thyroid hormones secreted by the thyroid and parathyroid primarily control metabolic rates and the regulation of serum calcium levels, respectively.
  • Thyroid hormones include calcitonin, somatostatin, and thyroid hormone.
  • the parathyroid secretes parathyroid hormone.
  • Disorders associated with hypothyroidism include goiter, myxedema, acute thyroiditis associated with bacterial infection, subacute thyroiditis associated with viral infection, autoimmune thyroiditis (Hashimoto's disease), and cretinism.
  • Disorders associated with hyperthyroidism include thyrotoxicosis and its various forms, Grave's disease, pretibial myxedema, toxic multinodular goiter, thyroid carcinoma, and Plummer's disease.
  • Disorders associated with hype ⁇ arathyroidism include Conn disease (chronic hypercalemia) leading to bone reso ⁇ tion and parathyroid hype ⁇ lasia.
  • Pancreatic hormones secreted by the pancreas regulate blood glucose levels by modulating the rates of carbohydrate, fat, and protein metaboUsm.
  • Pancreatic hormones include insuUn, glucagon, amyUn, ⁇ - aminobutyric acid, gastrin, somatostatin, and pancreatic polypeptide.
  • the principal disorder associated with pancreatic dysfunction is diabetes melUtus caused by insufficient insuUn activity. Diabetes melUtus is generally classified as either Type I (insulin-dependent, juvenile diabetes) or Type II (non- insuUn-dependent, adult diabetes). The treatment of both forms by insuUn replacement therapy is well known.
  • Diabetes melUtus often leads to acute complications such as hypoglycemia (insulin shock), 5 coma, diabetic ketoacidosis, lactic acidosis, and chronic compUcations leading to disorders of the eye, kidney, skin, bone, joint, cardiovascular system, nervous system, and to decreased resistance to infection.
  • hypoglycemia insulin shock
  • 5 coma diabetic ketoacidosis
  • lactic acidosis and chronic compUcations leading to disorders of the eye, kidney, skin, bone, joint, cardiovascular system, nervous system, and to decreased resistance to infection.
  • Growth factors are secreted proteins that mediate intercellular communication. UnUke hormones, which travel great distances via the circulatory system, most growth factors are primarily 5 local mediators that act on neighboring cells. Most growth factors contain a hydrophobic N-terminal signal peptide sequence which directs the growth factor into the secretory pathway. Most growth factors also undergo post-translational modifications within the secretory pathway. These modifications can include proteolysis, glycosylation, phosphorylation, and intramolecular disulfide bond formation. Once secreted, growth factors bind to specific receptors on the surfaces of neighboring o target cells, and the bound receptors trigger intracellular signal transduction pathways. These signal transduction pathways eUcit specific cellular responses in the target cells. These responses can include the modulation of gene expression and the stimulation or inhibition of cell division, cell differentiation, and cell motility.
  • the broadest class 5 includes the large polypeptide growth factors, which are wide-ranging in their effects. These factors include epidermal growth factor (EGF), fibroblast growth factor (FGF), transforming growth factor- ⁇ (TGF- ⁇ ), insulin-like growth factor (IGF), nerve growth factor (NGF), and platelet-derived growth factor (PDGF), each defining a family of numerous related factors.
  • EGF epidermal growth factor
  • FGF fibroblast growth factor
  • TGF- ⁇ transforming growth factor- ⁇
  • IGF insulin-like growth factor
  • NGF nerve growth factor
  • PDGF platelet-derived growth factor
  • the large polypeptide growth factors act as mitogens on diverse cell types to stimulate wound healing, o bone synthesis and remodeling, extracellular matrix synthesis, and proliferation of epithelial, epidermal, and connective tissues.
  • TGF- ⁇ , EGF, and FGF famines also function as inductive signals in the differentiation of embryonic tissue.
  • NGF functions specifically as a neurotrophic factor, promoting neuron
  • Another class of growth factors includes the hematopoietic growth factors, which are narrow in their target specificity. These factors stimulate the proUferation and differentiation of blood cells such as B-lymphocytes, T-lymphocytes, erythrocytes, platelets, eosinophils, basophils, neutrophils, macrophages, and their stem cell precursors. These factors include the colony-stimulating factors (G- CSF, M-CSF, GM-CSF, and CSF1-3), erythropoietin, and the cytokines. The cytokines are speciaUzed hematopoietic factors secreted by cells of the immune system and are discussed in detail below.
  • Growth factors play critical roles in neoplastic transformation of cells in vitro and in tumor progression in vivo. Overexpression of the large polypeptide growth factors promotes the proliferation and transformation of cells in culture. Inappropriate expression of these growth factors by tumor cells in vivo may contribute to tumor vascularization and metastasis. Inappropriate activity of hematopoietic growth factors can result in anemias, leukemias, and lymphomas. Moreover, growth factors are both structurally and functionally related to oncoproteins, the potentially cancer-causing products of proto-oncogenes. Certain FGF and PDGF family members are themselves homologous to oncoproteins, whereas receptors for some members of the EGF, NGF, and FGF families are encoded by proto-oncogenes.
  • Growth factors also affect the transcriptional regulation of both proto-oncogenes and oncosuppressor genes (Pimentel, E. (1994) Handbook of Growth Factors, CRC Press, Ann Arbor MI; McKay, I. and I. Leigh, eds. (1993) Growth Factors: A Practical Approach, Oxford University Press, New York NY; Habenicht, A., ed. (1990) Growth Factors. Differentiation Factors, and Cytokines. Springer- Verlag, New York NY).
  • NP/VM Small Peptide Factors - Neuropeptides and Vasomediators
  • neuropeptides and neuropeptide hormones such as bombesin, neuropeptide Y, neurotensin, neuromedin N, melanocortins, opioids, galanin, somatostatin, tachykinins, urotensin II and related peptides involved in smooth muscle stimulation, vasopressin, vasoactive intestinal peptide, and circulatory system-borne signaUng molecules such as angiotensin, complement, calcitonin, endotheUns, formyl-methionyl peptides, glucagon, cholecystokinin, gastrin, and many of the peptide hormones discussed above.
  • neuropeptides and neuropeptide hormones such as bombesin, neuropeptide Y, neurotensin, neuromedin N, melanocortins, opioids, galanin, somatostatin, tachykinins, urotensin II and related peptides involved in smooth
  • NP/VMs can transduce signals directly, modulate the activity or release of other neurotransmitters and hormones, and act as catalytic enzymes in signaUng cascades.
  • the effects of NP/VMs range from extremely brief to long-lasting. (Reviewed in Martin, CR. et al. (1985) Endocrine Physiology, Oxford University Press, New York NY, pp. 57-62.)
  • Cytokines function as growth and differentiation factors that act primarily on cells of the immune system such as B- and T-lymphocytes, monocytes, macrophages, and granulocytes. Like other signaUng molecules, cytokines bind to specific plasma membrane receptors and trigger o intracellular signal transduction pathways which alter gene expression patterns. There is considerable potential for the use of cytokines in the treatment of inflammation and immune system disorders.
  • Cytokine structure and function have been extensively characterized in vitro. Most cytokines are small polypeptides of about 30 kilodaltons or less. Over 50 cytokines have been identified from human and rodent sources. Examples of cytokine subfamiUes include the interferons (IFN- ⁇ , - ⁇ , and - 5 ⁇ ), the interleukins (ILl-ILl 3), the tumor necrosis factors (TNF- ⁇ and - ⁇ ), and the chemokines. Many cytokines have been produced using recombinant DNA techniques, and the activities of individual cytokines have been determined in vitro. These activities include regulation of leukocyte proliferation, differentiation, and motiUty.
  • cytokine activity may not reflect the full scope of that cytokine' s o activity in vivo.
  • Cytokines are not expressed individually in vivo but are instead expressed in combination with a multitude of other cytokines when the organism is challenged with a stimulus. Together, these cytokines collectively modulate the immune response in a manner appropriate for that particular stimulus. Therefore, the physiological activity of a cytokine is determined by the stimulus itself and by complex interactive networks among co-expressed cytokines which may demonstrate both 5 synergistic and antagonistic relationships.
  • Chemokines comprise a cytokine subfamily with over 30 members. (Reviewed in Wells, T. N.C. and M.C Peitsch (1997) J. Leukoc. Biol. 61:545-550.) Chemokines were initially identified as chemotactic proteins that recruit monocytes and macrophages to sites of inflammation. Recent evidence indicates that chemokines may also play key roles in hematopoiesis and HIV-1 infection. Chemokines 0 are small proteins which range from about 6-15 kilodaltons in molecular weight. Chemokines are further classified as C, CC, CXC, or CX 3 C based on the number and position of critical cysteine residues.
  • the CC chemokines for example, each contain a conserved motif consisting of two consecutive cysteines followed by two additional cysteines which occur downstream at 24- and 16- residue intervals, respectively (ExPASy PROSITE database, documents PS00472 and PDOC00434).
  • the presence and spacing of these four cysteine residues are highly conserved, whereas the intervening residues diverge significantly.
  • a conserved tyrosine located about 15 residues downstream of the cysteine doublet seems to be important for chemotactic activity.
  • Most of the human genes encoding CC chemokines are clustered on chromosome 17, although there are a few examples of CC chemokine genes that map elsewhere.
  • chemokines include lymphotactin (C chemokine); macrophage chemotactic and activating factor (MCAF/MCP-1; CC chemokine); platelet factor 4 and IL-8 (CXC chemokines); and fractalkine and neurofractin (CX 3 C chemokines).
  • SEQ ID NO:10 and SEQ ID NO:l 1 encode, for example, receptor molecules.
  • receptor describes proteins that specifically recognize other molecules.
  • the category is broad and includes proteins with a variety of functions.
  • the bulk of receptors are cell surface proteins which bind extracellular ligands and produce cellular responses in the areas of growth, differentiation, endocytosis, and immune response.
  • Other receptors faciUtate the selective transport of proteins out of the endoplasmic reticulum and locaUze enzymes to particular locations in the cell.
  • the term may also be apphed to proteins which act as receptors for Ugands with known or unknown chemical composition and which interact with other cellular components. For example, the steroid hormone receptors bind to and regulate transcription of DNA.
  • Regulatory proteins such as growth factors coordinately control these cellular processes and act as mediators in cell-cell signaUng pathways.
  • Growth factors are secreted proteins that bind to specific cell-surface receptors on target cells. The bound receptors trigger intracellular signal transduction pathways which activate various downstream effectors that regulate gene expression, cell division, cell differentiation, cell motility, and other cellular processes.
  • Cell surface receptors are typically integral plasma membrane proteins. These receptors recognize hormones such as catecholamines; peptide hormones; growth and differentiation factors; small peptide factors such as thyrotropin-releasing hormone; galanin, somatostatin, and tachykinins; and circulatory system-borne signaUng molecules.
  • Cell surface receptors on immune system cells recognize antigens, antibodies, and major histocompatibiUty complex (MHC)-bound peptides. Other cell surface receptors bind ligands to be internaUzed by the cell.
  • MHC major histocompatibiUty complex
  • LDL low density Upoproteins
  • transferrin glucose- or mannose-terminal glycoproteins, galactose-terminal glycoproteins, immunoglobulins, phosphovitellogenins, fibrin, proteinase-inhibitor complexes, plasminogen activators, and thrombospondin
  • phosphovitellogenins fibrin
  • proteinase-inhibitor complexes proteinase-inhibitor complexes
  • plasminogen activators and thrombospondin
  • growth factor receptors including receptors for epidermal growth factor, 5 platelet-derived growth factor, fibroblast growth factor, as well as the growth modulator ⁇ -thrombin, contain intrinsic protein kinase activities. When growth factor binds to the receptor, it triggers the autophosphorylation of a serine, threonine, or tyrosine residue on the receptor. These phosphorylated sites are recognition sites for the binding of other cytoplasmic signaUng proteins. These proteins participate in signaUng pathways that eventually link the initial receptor activation at the cell surface to 0 the activation of a specific intracellular target molecule.
  • SH2 domains and SH3 domains are found in phosphoUpase C- ⁇ , PI-3-K p85 regulatory subunit, Ras-GTPase activating protein, and pp ⁇ O 0 (Lowenstein, E.J. et al. (1992) Cell 70:431-442).
  • the cytokine family of receptors share a different common binding domain and include transmembrane 5 receptors for growth hormone (GH), interleukins, erythropoietin, and prolactia
  • receptors and second messenger-binding proteins have intrinsic serine/threonine protein kinase activity. These include activin/TGF- ⁇ /BMP-superfamily receptors, calcium- and diacylglycerol- activated/phosphoUpid-dependant protein kinase (PK-C), and RNA-dependant protein kinase (PK-R).
  • PKI activin/TGF- ⁇ /BMP-superfamily receptors
  • PK-C calcium- and diacylglycerol- activated/phosphoUpid-dependant protein kinase
  • PK-R RNA-dependant protein kinase
  • serine/threonine protein kinases including nematode Twitchin, have fibronectin-Uke, o immunoglobuUn C2-Uke domains.
  • G-protein coupled receptors are integral membrane proteins characterized by the presence of seven hydrophobic transmembrane domains which span the plasma membrane and form a bundle of antiparallel alpha ( ⁇ ) heUces. These proteins range in size from under 400 to over 1000 5 amino acids (Strosberg, AD. (1991) Eur. J. Biochem. 196:1-10; CoughUn, S.R. (1994) Curr. Opin. Cell Biol. 6:191-197).
  • the amino-terminus of the GPCR is extracellular, of variable length and often glycosylated; the carboxy-terminus is cytoplasmic and generally phosphorylated. Extracellular loops of the GPCR alternate with intracellular loops and Unk the transmembrane domains.
  • the most conserved domains of GPCRs are the transmembrane domains and the first two cytoplasmic loops.
  • the o transmembrane domains account for structural and functional features of the receptor.
  • the bundle of ⁇ heUces forms a binding pocket.
  • the extracellular N-terminal segment or one or more of the three extracellular loops may also participate in Ugand binding.
  • Ligand binding activates the receptor by inducing a conformational change in intracellular portions of the receptor.
  • the activated receptor interacts with an intracellular heterotrimeric guanine nucleotide binding (G) protein complex which mediates further intracellular signaling activities, generally the production of second messengers such as cycUc AMP (cAMP), phosphoUpase C, inositol triphosphate, or interactions with ion channel proteins (Baldwin, J.M. (1994) Curr. Opin. Cell Biol. 6:180-190).
  • G guanine nucleotide binding
  • GPCRs include those for acetylchoUne, adenosine, epinephrine and norepinephrine, bombesin, 5 bradykinin, chemokines, dopamine, endotheUn, ⁇ -aminobutyric acid (GABA), folUcle-stimulating hormone (FSH), glutamate, gonadotropin-releasing hormone (GnRH), hepatocyte growth factor, histamine, leukotrienes, melanocortins, neuropeptide Y, opioid peptides, opsins, prostanoids, serotonin, somatostatin, tachykinins, thrombin, thyrotropin-releasing hormone (TRH), vasoactive intestinal polypeptide family, vasopressin and oxytocin, and o ⁇ han receptors.
  • GABA ⁇ -aminobutyric acid
  • FSH folUcle-stimulating hormone
  • retinitis pigmentosa may arise from mutations in the rhodopsin gene.
  • Rhodopsin is the retinal photoreceptor which is located within the discs of the eye rod cell.
  • Parma, J. et al. (1993, Nature 365:649-651) report that somatic activating mutations in the thyrotropin receptor cause hyperfunctioning thyroid adenomas and suggest 5 that certain GPCRs susceptible to constitutive activation may behave as protooncogenes.
  • Nuclear receptors bind small molecules such as hormones or second messengers, leading to increased receptor-binding affinity to specific chromosomal DNA elements. In addition the affinity for other nuclear proteins may also be altered. Such binding and protein-protein interactions may regulate o and modulate gene expression. Examples of such receptors include the steroid hormone receptors family, the retinoic acid receptors family, and the thyroid hormone receptors family. Ligand-Gated Receptor Ion Channels
  • Ligand-gated receptor ion channels fall into two categories.
  • the first category extracellular Ugand-gated receptor ion channels (ELGs), rapidly transduce neurotransmitter-binding events into 5 electrical signals, such as fast synaptic neurotransmission. ELG function is regulated by post- translational modification.
  • the second category intracellular ligand-gated receptor ion channels (ILGs), are activated by many intracellular second messengers and do not require post-translational modifications) to effect a channel-opening response.
  • ELGs depolarize excitable cells to the threshold of action potential generation. In non-excitable o cells, ELGs permit a Umited calcium ion-influx during the presence of agonist.
  • ELGs include channels directly gated by neurotransmitters such as acetylchoUne, L-glutamate, glycine, ATP, serotonin, GABA, and histamine.
  • ELG genes encode proteins having strong structural and functional similarities. ILGs are encoded by distinct and unrelated gene famiUes and include receptors for cAMP, cGMP, calcium ions, ATP, and metaboUtes of arachidonic acid. Macrophage Scavenger Receptors
  • Macrophage scavenger receptors with broad ligand specificity may participate in the binding of low density Upoproteins (LDL) and foreign antigens.
  • Scavenger receptors types I and II are trimeric membrane proteins with each subunit containing a small N-terminal intracellular domain, a 5 transmembrane domain, a large extracellular domain, and a C-terminal cysteine-rich domain.
  • the extracellular domain contains a short spacer domain, an ⁇ -heUcal coiled-coil domain, and a triple heUcal collagenous domain.
  • T cells play a dual role in the immune system as effectors and regulators, coupUng antigen 5 recognition with the transmission of signals that induce cell death in infected cells and stimulate proUferation of other immune cells.
  • TCR T cell receptor
  • MHC major histocompatibiUty molecule
  • Both TCR subunits have an extracellular domain containing both variable and constant regions, a transmembrane domain that traverses the membrane once, and a short intracellular domain (Saito, H. et al. (1984) Nature 309:757-762).
  • the genes for the TCR subunits are constructed through somatic rearrangement of different gene segments. Interaction of antigen in the proper MHC context 5 with the TCR initiates signaling cascades that induce the proUferation, maturation, and function of cellular components of the immune system (Weiss, A. (1991) Annu. Rev. Genet. 25:487-510).
  • SEQ ID NO:12, SEQ ID NO:13, SEQ ID NO:14, SEQ ID NO:15, SEQ ID NO:16, SEQ ID NO: 17, and SEQ ID NO: 18 encode, for example, intracellular signaling molecules.
  • Intracellular signaling is the general process by which cells respond to extracellular signals (hormones, neurotransmitters, growth and differentiation factors, etc.) through a cascade of biochemical reactions that begins with the binding of a signaling molecule to a cell membrane receptor and ends with the activation of an intracellular target molecule.
  • Intermediate steps in the process involve the activation of various cytoplasmic proteins by phosphorylation via protein kinases, 5 and their deactivation by protein phosphatases, and the eventual translocation of some of these activated proteins to the cell nucleus where the transcription of specific genes is triggered.
  • the intracellular signaling process regulates all types of cell functions including cell proliferation, cell differentiation, and gene transcription, and involves a diversity of molecules including protein kinases and phosphatases, and second messenger molecules, such as cyclic nucleotides, calcium-calmodulin, o inositol, and various mitogens, that regulate protein phosphorylation.
  • second messenger molecules such as cyclic nucleotides, calcium-calmodulin, o inositol, and various mitogens, that regulate protein phosphorylation.
  • Protein kinases and phosphatases play a key role in the intracellular signaling process by controlling the phosphorylation and activation of various signaling proteins.
  • the high energy phosphate for this reaction is generally transferred from the adenosine triphosphate molecule (ATP) to 5 a particular protein by a protein kinase and removed from that protein by a protein phosphatase.
  • ATP adenosine triphosphate molecule
  • Protein kinases are roughly divided into two groups: those that phosphorylate tyrosine residues (protein tyrosine kinases, PTK) and those that phosphorylate serine or threonine residues (serine/threonine kinases, STK).
  • a few protein kinases have dual specificity for serine/threonine and tyrosine residues. Almost all kinases contain a conserved 250-300 amino acid catalytic domain o containing specific residues and sequence motifs characteristic of the kinase family (Hardie, G. and S.
  • STKs include the second messenger dependent protein kinases such as the cyclic- AMP dependent protein kinases (PKA), involved in mediating hormone-induced cellular responses; calcium-calmodulin (CaM) dependent protein kinases, involved in regulation of smooth muscle 5 contraction, glycogen breakdown, and neurofransmission; and the mitogen-activated protein kinases
  • PKA cyclic- AMP dependent protein kinases
  • CaM calcium-calmodulin dependent protein kinases
  • MAP MAP which mediate signal transduction from the cell surface to the nucleus via phosphorylation cascades.
  • Altered PKA expression is implicated in a variety of disorders and diseases including cancer, thyroid disorders, diabetes, atherosclerosis, and cardiovascular disease (Isselbacher, KJ. et al. (1994) Harrison's Principles of Internal Medicine McGraw-Hill, New York NY, pp. 416-431, 1887).
  • o PTKs are divided into transmembrane, receptor PTKs and nontransmembrane, non-receptor
  • Transmembrane PTKs are receptors for most growth factors.
  • Non-receptor PTKs lack transmembrane regions and, instead, form complexes with the intracellular regions of cell surface receptors.
  • Receptors that function through non-receptor PTKs include those for cytokines and hormones (growth hormone and prolactin) and antigen-specific receptors on T and B lymphocytes. 5 Many of these PTKs were first identified as the products of mutant oncogenes in cancer cells in which their activation was no longer subject to normal cellular controls.
  • HPK histidine protein kinase family
  • a histidine residue in the N-terminal half of the molecule (region I) is an autophosphorylation site.
  • Three additional motifs located in the C-terminal half of the molecule 0 include an invariant asparagine residue in region II and two glycine-rich loops characteristic of nucleotide binding domains in regions III and IV.
  • Recently a branched chain alpha-ketoacid dehydrogenase kinase has been found with characteristics of HPK in rat (Davie, supra).
  • the two principal categories of protein phosphatases 5 are the protein (serine/threonine) phosphatases (PPs) and the protein tyrosine phosphatases (PTPs).
  • PPs dephosphorylate phosphoserine/threonine residues and are important regulators of many cAMP-mediated hormone responses (Cohen, P. (1989) Annu. Rev. Biochem. 58:453-508).
  • PTPs reverse the effects of protein tyrosine kinases and play a significant role in cell cycle and cell signaling processes (Charbonneau, supra).
  • PTPs may prevent or reverse cell transformation and the growth of various cancers by controlling the levels of tyrosine phosphorylation in cells. This hypothesis is supported by studies showing that overexpression of PTPs can suppress transformation in cells, and that specific inhibition of PTPs can enhance cell transformation (Charbonneau, supra). 5 Phospholipid and Inositol-Phosphate Signaling
  • Inositol phospholipids are involved in an intracellular signaling pathway that begins with binding of a signaling molecule to a G-protein linked receptor in the plasma membrane. This leads to the phosphorylation of phosphatidylinositol (PI) residues on the inner side of the plasma membrane to the biphosphate state (PIP j ) by inositol kinases. Simultaneously, the G- o protein Unked receptor binding stimulates a trimeric G-protein which in turn activates a phosphoinositide-specific phospholipase C- ⁇ .
  • PI phosphatidylinositol
  • IP 3 inositol triphosphate
  • diacylglycerol acts as mediators for separate signaling events.
  • IP 3 diffuses through the plasma membrane to induce calcium release from the endoplasmic reticulum (ER), while diacylglycerol remains in the membrane and helps activate 5 protein kinase C, an STK that phosphorylates selected proteins in the target cell.
  • the calcium response initiated by IP 3 is terminated by the dephosphorylation of IP 3 by specific inositol phosphatases.
  • Cyclic nucleotides function as intracellular second messengers to transduce a variety of extracellular signals including hormones, light, and neurotransmitters.
  • cyclic- AMP dependent protein kinases PKA
  • PKA cyclic- AMP dependent protein kinases
  • adenylyl cyclase which synthesizes cAMP from AMP, is activated to increase cAMP levels in muscle by binding of adrenaUne to ⁇ -andrenergic receptors, while activation 5 of guanylate cyclase and increased cGMP levels in photoreceptors leads to reopening of the
  • PDEs cAMP and cGMP-specific phosphodiesterases
  • PDE inhibitors have been found to be particularly useful in treating various clinical disorders.
  • Rolipram a specific inhibitor of PDE4
  • TheophyUine is a nonspecific PDE inhibitor used in the treatment of bronchial asthma and other respiratory diseases (Banner, K.H. and CP. Page (1995) Eur. Respir. J. 8:996-1000).
  • G-proteins are critical mediators of signal transduction between a particular class of extracellular receptors, the G-protein coupled receptors (GPCR), and 0 intracellular second messengers such as cAMP and Ca 2+ .
  • G-proteins are linked to the cytosoUc side of a GPCR such that activation of the GPCR by ligand binding stimulates binding of the G-protein to GTP, inducing an "active" state in the G-protein. In the active state, the G-protein acts as a signal to trigger other events in the cell such as the increase of cAMP levels or the release of Ca 2+ into the cytosol from the ER, which, in turn, regulate phosphorylation and activation of other intracellular 5 proteins.
  • the ⁇ and ⁇ subunits form a tight complex that anchors the protein to the inner side of the plasma membrane.
  • the ⁇ subunits also known as G- ⁇ proteins or ⁇ transducins, contain seven tandem repeats of the WD-repeat sequence motif, a motif found in many proteins with regulatory functions. Mutations and variant expression of ⁇ fransducin proteins are o linked with various disorders (Neer, E. J. et al. (1994) Nature 371 :297-300; Margottin, F. et al. (1998)
  • LMW GTP-proteins are GTPases which regulate cell growth, cell cycle control, protein secretion, and intracellular vesicle interaction. They consist of single polypeptides which, Uke the ⁇ subunit of the heterotrimeric G-proteins, are able to bind and hydrolyze GTP, thus cycling between an 5 inactive and an active state. At least sixty members of the LMW G-protein superfamily have been identified and are currently grouped into the six subfamilies of ras, rho, arf, sari, ran, and rab. Activated ras genes were initially found in human cancers, and subsequent studies confirmed that ras function is critical in determining whether cells continue to grow or become differentiated. Other members of the LMW G-protein superfamily have roles in signal transduction that vary with the o function of the activated genes and the locations of the G-proteins.
  • Guanine nucleotide exchange factors regulate the activities of LMW G-proteins by determining whether GTP or GDP is bound.
  • GTPase-activating protein GAP
  • GTP-ras binds to GTP-ras and induces it to hydrolyze GTP to GDP.
  • GNRP guanine nucleotide releasing protein
  • RGS G-protein signaling
  • RGS family members are related structurally through similarities in an approximately 120 amino acid region termed the RGS domain and functionally by their abiUty to inhibit the interleukin (cytokine) induction of MAP kinase 0 in cultured mammalian 293T cells (Druey, supra).
  • cytokine interleukin
  • Ca +2 is another second messenger molecule that is even more widely used as an intracellular mediator than cAMP.
  • Ca 2+ directly activates regulatory enzymes, such as protein kinase C, which trigger signal transduction pathways.
  • Ca 2+ also binds to specific Ca 2+ -binding proteins (CBPs) 5 such as calmoduUn (CaM) which then activate multiple target proteins in the cell including enzymes, membrane transport pumps, and ion channels.
  • CBPs Ca 2+ -binding proteins
  • CaM interactions are involved in a multitude of cellular processes including, but not limited to, gene regulation, DNA synthesis, cell cycle progression, mitosis, cytokinesis, cytoskeletal organization, muscle contraction, signal transduction, ion homeostasis, exocytosis, and metabolic regulation (Celio, M.R. et al. (1996) Guidebook to 0 Calcium-binding Proteins, Oxford University Press, Oxford, UK, pp. 15-20).
  • CBPs can serve as a storage depot for Ca 2+ in an inactive state.
  • Calsequestrin is one such CBP that is expressed in isoforms specific to cardiac muscle and skeletal muscle. It is suggested that calsequestrin binds Ca 2+ in a rapidly exchangeable state that is released during Ca 2+ -signaling conditions (Celio, M.R. et al. (1996) Guidebook to Calcium-binding Proteins, Oxford University Press, New York NY, pp. 222- 5 224).
  • Cell division is the fundamental process by which all Uving things grow and reproduce. In most organisms, the cell cycle consists of three principle steps; inte ⁇ hase, mitosis, and cytokinesis. Inte ⁇ hase, involves preparations for cell division, replication of the DNA and production of essential o proteins. In mitosis, the nuclear material is divided and separates to opposite sides of the cell.
  • Cytokinesis is the final division and fission of the cell cytoplasm to produce the daughter cells.
  • Cyclins act by binding to and activating a group of cycUn-dependent protein kinases (Cdks) which then phosphorylate and activate selected proteins 5 involved in the mitotic process.
  • Cdks cycUn-dependent protein kinases
  • Two principle types are mitotic cycUn, or cyclin B, which controls entry of the cell into mitosis, and Gl cyclin, which controls events that drive the cell out of mitosis.
  • cyclin B which controls entry of the cell into mitosis
  • Gl cyclin which controls events that drive the cell out of mitosis.
  • Ceretain proteins in intracellular signaling pathways serve to link or cluster other proteins o involved in the signaling cascade.
  • a conserved protein domain called the PDZ domain has been identified in various membrane-associated signaling proteins. This domain has been implicated in receptor and ion channel clustering and in the targeting of multiprotein signaling complexes to specialized functional regions of the cytosoUc face of the plasma membrane. (For a review of PDZ domain-containing proteins, see Ponting, CP. et al.
  • PDZ domains are found in the eukaryotic MAGUK (membrane-associated guanylate kinase) protein family, members of which bind to the intracellular domains of receptors and channels.
  • MAGUK membrane-associated guanylate kinase
  • PDZ domains are also found in diverse membrane-localized proteins such as protein tyrosine phosphatases, serine/threonine kinases, G-protein cofactors, and synapse-associated proteins 5 such as syntrophins and neuronal nitric oxide synthase (nNOS).
  • nNOS neuronal nitric oxide synthase
  • Membrane Transport Molecules o
  • the plasma membrane acts as a barrier to most molecules. Transport between the cytoplasm and the extracellular environment, and between the cytoplasm and lumenal spaces of cellular organelles requires specific transport proteins.
  • Each transport protein carries a particular class of molecule, such as ions, sugars, or amino acids, and often is specific to a certain molecular species of the class.
  • a variety of human inherited diseases are caused by a mutation in a transport protein. For 5 example, cystinuria is an inherited disease that results from the inability to transport cystine, the disulfide-linked dimer of cysteine, from the urine into the blood. Accumulation of cystine in the urine leads to the formation of cystine stones in the kidneys.
  • Transport proteins are multi-pass transmembrane proteins, which either actively transport molecules across the membrane or passively allow them to cross. Active transport involves o directional pumping of a solute across the membrane, usually against an electrochemical gradient.
  • Active transport is tightly coupled to a source of metabolic energy, such as ATP hydrolysis or an elecfrochemically favorable ion gradient.
  • Passive transport involves the movement of a solute down its electrochemical gradient.
  • Transport proteins can be further classified as either carrier proteins or channel proteins.
  • Carrier proteins which can function in active or passive transport, bind to a specific 5 solute to be transported and undergo a conf ormational change which transfers the bound solute across the membrane.
  • Channel proteins which only function in passive transport, form hydrophilic pores across the membrane. When the pores open, specific solutes, such as inorganic ions, pass through the membrane and down the electrochemical gradient of the solute.
  • Carrier proteins which transport a single solute from one side of the membrane to the other 0 are called uniporters.
  • coupled transporters link the transfer of one solute with simultaneous or sequential transfer of a second solute, either in the same direction (symport) or in the opposite direction (antiport).
  • intestinal and kidney epithelium contains a variety of symporter systems driven by the sodium gradient that exists across the plasma membrane. Sodium moves into the cell down its electrochemical gradient and brings the solute into the cell with it. The 5 sodium gradient that provides the driving force for solute uptake is maintained by the ubiquitous Na7K + ATPase.
  • Sodium-coupled transporters include the mammaUan glucose transporter (SGLT1), iodide transporter (NIS), and multivitamin transporter (SMVT). Al three transporters have twelve putative transmembrane segments, extracellular glycosylation sites, and cytoplasmically-oriented N- and C-termini. NIS plays a crucial role in the evaluation, diagnosis, and treatment of various thyroid 5 pathologies because it is the molecular basis for radioiodide thyroid-imaging techniques and for specific targeting of radioisotopes to the thyroid gland (Levy, O. et al. (1997) Proc. Natl. Acad. Sci. USA 94:5568-5573).
  • SMVT is expressed in the intestinal mucosa, kidney, and placenta, and is implicated in the transport of the water-soluble vitamins, e.g., biotin and pantothenate (Prasad, P.D. et al. (1998) J. Biol. Chem. 273:7501-7506).
  • o Transporters play a major role in the regulation of pH, excretion of drugs, and the cellular
  • Monocarboxylate anion transporters are proton-coupled symporters with a broad substrate specificity that includes L-lactate, pyruvate, and the ketone bodies acetate, acetoacetate, and beta-hydroxybutyrate. At least seven isoforms have been identified to date. The isoforms are predicted to have twelve transmembrane (TM) heUcal domains with a large intracellular loop between TM6 and 5 TM7, and play a critical role in maintaining intracellular pH by removing the protons that are produced stoichiometrically with lactate during glycolysis.
  • TM transmembrane
  • H(+)-monocarboxylate transporter is that of the erythrocyte membrane, which transports L-lactate and a wide range of other aUphatic monocarboxylates.
  • Other cells possess H(+)-Unked monocarboxylate transporters with differing substrate and inhibitor selectivities.
  • cardiac muscle and tumor cells have o transporters that differ in their K m values for certain substrates, including stereoselectivity for L- over
  • D-lactate D-lactate, and in their sensitivity to inhibitors.
  • Na(+)-monocarboxylate cotransporters on the luminal surface of intestinal and kidney epitheUa, which allow the uptake of lactate, pyruvate, and ketone bodies in these tissues.
  • transporters for organic cations and organic anions in organs including the kidney, intestine and liver are selective for hydrophobic, charged molecules with electron-attracting side groups.
  • Organic cation transporters such as the ammonium transporter, mediate the secretion of a variety of drugs and endogenous metaboUtes, and contribute to the maintenance of intercellular pH.
  • Am. J. Physiol. 264:C761-C782 mediate the secretion of a variety of drugs and endogenous metaboUtes, and contribute to the maintenance of intercellular pH.
  • AP. Halestrap (1993) Am. J. Physiol. 264:C761-C782; Price, N.T. et al. (1998) Biochem. J. 329:321-328; and Martinelle, K. and I. Haggstrom (1993) J. Biotechnol. 30: 339-350.
  • ATP-binding cassette The largest and most diverse family of transport proteins known is the ATP-binding cassette
  • ABC transporters can transport substances that differ markedly in chemical structure and size, ranging from small molecules such as ions, sugars, amino acids, peptides, and phosphoUpids, to Upopeptides, large proteins, and complex hydrophobic drugs.
  • ABC proteins consist of four modules: two nucleotide-binding domains (NBD), which hydrolyze ATP to supply the energy required for transport, and two membrane-spanning domains (MSD), each containing six putative transmembrane segments. These four modules may be encoded by a single gene, as is the case for the cystic fibrosis transmembrane regulator (CFTR), or by separate genes.
  • NBD nucleotide-binding domains
  • MSD membrane-spanning domains
  • each gene product contains a single NBD and MSD. These "half-molecules" form 5 homo- and heterodimers, such as Tapl and Tap2, the endoplasmic reticulum-based major histocompatibiUty (MHC) peptide transport system.
  • MHC major histocompatibiUty
  • CFTR cystic fibrosis
  • ALDP adrenoleukodystrophy
  • ALDP adrenoleukodysfrophy protein
  • Zellweger syndrome peroxisomal membrane protein-70, PMP70
  • hyperinsuUnemic hypoglycemia 0 sulfonylurea receptor, SUR.
  • MDR multidrug resistance
  • ABC transporter in human cancer cells makes the cells resistant to a variety of cytotoxic drugs used in chemotherapy (TagUght, D. and S. MichaeUs (1998) Meth. Enzymol. 292:131-163).
  • Fatty acid transport protein an integral membrane protein with four transmembrane segments, is expressed in tissues exhibiting high levels of plasma membrane fatty acid flux, such as muscle, heart, and adipose. Expression of FATP is upregulated in 3T3-L1 cells during adipose conversion, and expression in COS7 fibroblasts elevates uptake of long-chain fatty acids (Hui, T.Y. et al. (1998) J. o Biol. Chem. 273 :27420-27429). Ion Channels
  • the electrical potential of a cell is generated and maintained by controlUng the movement of ions across the plasma membrane.
  • the movement of ions requires ion channels, which form an ion- selective pore within the membrane.
  • ion channels There are two basic types of ion channels, ion transporters and 5 gated ion channels.
  • Ion transporters utilize the energy obtained from ATP hydrolysis to actively transport an ion against the ion's concentration gradient.
  • Gated ion channels allow passive flow of an ion down the ion's electrochemical gradient under restricted conditions.
  • these types of ion channels generate, maintain, and utitize an electrochemical gradient that is used in 1) electrical impulse conduction down the axon of a nerve cell, 2) transport of molecules into cells against concentration o gradients, 3) initiation of muscle contraction, and 4) endocrine cell secretion.
  • Ion transporters generate and maintain the resting electrical potential of a cell. UtiUzing the energy derived from ATP hydrolysis, they transport ions against the ion's concentration gradient. These transmembrane ATPases are divided into three famiUes.
  • the phosphorylated (P) class ion transporters including Na + -K + ATPase, Ca 2+ -ATPase, and H + -ATPase, are activated by a phosphorylation event.
  • P-class ion transporters are responsible for maintaining resting potential distributions such that cytosoUc concentrations of Na + and Ca 2+ are low and cytosoUc concentration of K + is high.
  • the vacuolar (V) class of ion transporters includes H + pumps on intracellular organelles, such as lysosomes and Golgi. V-class ion transporters are responsible for generating the low pH within 5 the lumen of these organelles that is required for function.
  • the coupUng factor (F) class consists of H + pumps in the mitochondria. F-class ion transporters utiUze a proton gradient to generate ATP from ADP and inorganic phosphate (PJ.
  • the resting potential of the cell is utiUzed in many processes involving carrier proteins and gated ion channels.
  • Carrier proteins utilize the resting potential to transport molecules into and out of 0 the cell.
  • Amino acid and glucose transport into many cells is Unked to sodium ion co-transport
  • Ion channels share common structural and mechanistic themes.
  • the channel consists of four or 5 five subunits or protein monomers that are arranged like a barrel in the plasma membrane. Each subunit typically consists of six potential transmembrane segments (SI, S2, S3, S4, S5, and S6).
  • the center of the barrel forms a pore lined by ⁇ -helices or ⁇ -strands.
  • the side chains of the amino acid residues comprising the ⁇ -heUces or ⁇ -strands e stabUsh the charge (cation or anion) selectivity of the channel.
  • the degree of selectivity, or what specific ions are allowed to pass through the channel o depends on the diameter of the narrowest part of the pore.
  • Gated ion channels control ion flow by regulating the opening and closing of pores. These channels are categorized according to the manner of regulating the gating function. Mechanically-gated channels open pores in response to mechanical stress, voltage-gated channels open pores in response to changes in membrane potential, and Ugand-gated channels open pores in the presence of a specific ion, 5 nucleotide, or neurotransmitter.
  • Voltage-gated Na + and K + channels are necessary for the function of electrically excitable cells, such as nerve and muscle cells. Action potentials, which lead to neurotransmitter release and muscle contraction, arise from large, transient changes in the permeabiUty of the membrane to Na + and K + ions. Depolarization of the membrane beyond the threshold level opens voltage-gated Na + channels. Sodium o ions flow into the cell, further depolarizing the membrane and opening more voltage-gated Na + channels, which propagates the depolarization down the length of the cell. Depolarization also opens voltage-gated potassium channels. Consequently, potassium ions flow outward, which leads to repolarization of the membrane.
  • Voltage-gated channels utilize charged residues in the fourth transmembrane segment (S4) to sense voltage change.
  • the open state lasts only about 1 millisecond, at which time the channel spontaneously converts into an inactive state that cannot be opened irrespective of the membrane potential.
  • Inactivation is mediated by the channel's N-terminus, which acts as a plug that closes the pore.
  • the transition from an inactive to a closed state requires a return to resting potential.
  • Voltage-gated Na + channels are heterotrimeric complexes composed of a 260 kDa pore forming ⁇ subunit that associates with two smaller auxiliary subunits, ⁇ l and ⁇ 2.
  • the ⁇ 2 subunit is an integral membrane glycoprotein that contains an extracellular Ig domain, and its association with ⁇ and ⁇ l subunits correlates with increased functional expression of the channel, a change in its gating properties, and an increase in whole cell capacitance due to an increase in membrane surface area.
  • Voltage-gated Ca 2+ channels are involved in presynaptic neurotransmitter release, and heart and skeletal muscle contraction.
  • the voltage-gated Ca 2+ channels from skeletal muscle (L-type) and brain (N-type) have been purified, and though their functions differ dramatically, they have similar subunit compositions.
  • the channels are composed of three subunits.
  • the a ⁇ subunit forms the membrane pore and voltage sensor, while the ⁇ 2 ⁇ and ⁇ subunits modulate the voltage-dependence, gating properties, and the current amplitude of the channel.
  • These subunits are encoded by at least six ⁇ 1( one 0 2 8, and four ⁇ genes.
  • a fourth subunit, ⁇ has been identified in skeletal muscle. (Walker, D. et al.
  • Chloride channels are necessary in endocrine secretion and in regulation of cytosoUc and organelle pH.
  • Cl " enters the cell across a basolateral membrane through an Na + , K7C1 " cotransporter, accumulating in the cell above its electrochemical equilibrium concentration. Secretion of Cl " from the apical surface, in response to hormonal stimulation, leads to flow of Na + and water into the secretory lumen.
  • the cystic fibrosis transmembrane conductance regulator is a chloride channel encoded by the gene for cystic fibrosis, a common fatal genetic disorder in humans. Loss of CFTR function decreases transepitheUal water secretion and, as a result, the layers of mucus that coat the respiratory tree, pancreatic ducts, and intestine are dehydrated and difficult to clear. The resulting blockage of these sites leads to pancreatic insufficiency, "meconium ileus", and devastating "chronic obstructive pulmonary disease” (A-Awqati, Q. et al. (1992) J. Exp. Biol. 172:245-266).
  • H + - ATPase pumps that generate transmembrane pH and electrochemical differences by moving protons from the cytosol to the organelle lumen. If the membrane of the organelle is permeable to other ions, then the electrochemical gradient can be abrogated without affecting the pH differential. In fact, removal of the electrochemical barrier allows more H + to be pumped across the membrane, increasing the pH differential.
  • Cl " is the sole counterion of H + translocation in a number of organelles, including chromaffin granules, Golgi vesicles, lysosomes, and endosomes.
  • Functions that require a low vacuolar pH include uptake of small molecules such as biogenic amines in chromaffin granules, processing of vacuolar constituents such as pro-hormones by proteolytic enzymes, and protein degradation in lysosomes (A-Awqati, supra).
  • Ligand-gated channels open their pores when an extracellular or intracellular mediator binds to 5 the channel.
  • Neurotransmitter-gated channels are channels that open when a neurotransmitter binds to their extracellular domain. These channels exist in the postsynaptic membrane of nerve or muscle cells.
  • Chloride channels open in response to inhibitory neurotransmitters, such as ⁇ -aminobutyric acid (GABA) and glycine, leading to hype ⁇ olarization of the membrane and the subsequent generation of an action potential.
  • GABA ⁇ -aminobutyric acid
  • Ligand-gated channels can be regulated by intracellular second messengers. Calcium-activated K + channels are gated by internal calcium ions. In nerve cells, an influx of calcium during 5 depolarization opens K + channels to modulate the magnitude of the action potential (Ishi, T.M. et al. (1997) Proc. Nail. Acad. Sci. USA 94:11651-11656). Cyclic micleotide-gated (CNG) channels are gated by cytosoUc cyclic nucleotides. The best examples of these are the cAMP-gated Na + channels involved in olfaction and the cGMP-gated cation channels involved in vision. Both systems involve ligand-mediated activation of a G-protein coupled receptor which then alters the level of cyclic o nucleotide within the cell.
  • Ion channels are expressed in a number of tissues where they are impUcated in a variety of processes.
  • CNG channels while abundantly expressed in photoreceptor and olfactory sensory cells, are also found in kidney, lung, pineal, retinal gangUon cells, testis, aorta, and brain.
  • Calcium-activated K + channels may be responsible for the vasodilatory effects of bradykinin in the kidney and for shunting 5 excess K + from brain capillary endotheUal cells into the blood. They are also impUcated in repolarizing granulocytes after agonist-stimulated depolarization (Ishi, supra). Ion channels have been the target for many drug therapies.
  • Neurotransmitter-gated channels have been targeted in therapies for treatment of insomnia, anxiety, depression, and schizophrenia.
  • Voltage-gated channels have been targeted in therapies for arrhythmia, ischemic stroke, head trauma, and neurodegenerative disease (Taylor, CP. o and L.S. Narasimhan (1997) Adv. Pharmacol. 39:47-98).
  • SEQ ID NO:34 encodes, for example, a protein modification and maintenance molecule.
  • the cellular processes regulating modification and maintenance of protein molecules o coordinate their conformation, stabilization, and degradation. Each of these processes is mediated by key enzymes or proteins such as proteases, protease inhibitors, transferases, isomerases, and molecular chaperones.
  • Proteases cleave proteins and peptides at the peptide bond that forms the backbone of the 5 peptide and protein chain.
  • Proteolytic processing is essential to cell growth, differentiation, remodeling, and homeostasis as well as inflammation and immune response. Typical protein half- lives range from hours to a few days, so that within all living cells, precursor proteins are being cleaved to their active form, signal sequences proteolytically removed from targeted proteins, and aged or defective proteins degraded by proteolysis.
  • Proteases function in bacterial, parasitic, and viral o invasion and replication within a host.
  • Four principal categories of mammalian proteases have been identified based on active site structure, mechanism of action, and overall three-dimensional structure. (Beynon, R.J. and J.S. Bond (1994) Proteolytic Enzymes: A Practical Approach, Oxford University Press, New York NY, pp. 1-5).
  • SPs serine proteases
  • the serine proteases have a serine residue, usually within a conserved sequence, in an 5 active site composed of the serine, an aspartate, and a histidine residue.
  • SPs include the digestive enzymes trypsin and chymotrypsin, components of the complement cascade and the blood-clotting cascade, and enzymes that control extracellular protein degradation.
  • the main SP sub-families are trypases, which cleave after arginine or lysine; aspartases, which cleave after aspartate; chymases, which cleave after phenylalanine or leucine; metases, which cleavage after methionine; and serases o which cleave after serine.
  • Enterokinase the initiator of intestinal digestion, is a serine protease found in the intestinal brush border, where it cleaves the acidic propeptide from trypsinogen to yield active trypsin (Kitamoto, Y. et al. (1994) Proc. Natl. Acad. Sci. USA 91:7588-7592).
  • Prolylcarboxypeptidase a lysosomal serine peptidase that cleaves peptides such as angiotensin II and III and [des- Ag9] bradykinin, shares sequence homology with members of both the serine carboxypeptidase and prolylendopeptidase families (Tan, F. et al. (1993) J. Biol. Chem. 268:16631- 16638).
  • Cysteine proteases have a cysteine as the major catalytic residue at an active site where catalysis proceeds via an intermediate thiol ester and is facilitated by adjacent histidine and aspartic 5 acid residues.
  • CPs are involved in diverse cellular processes ranging from the processing of precursor proteins to intracellular degradation. Mammalian CPs include lysosomal cathepsins and cytosoUc calcium activated proteases, calpains.
  • CPs are produced by monocytes, macrophages and other cells of the immune system which migrate to sites of inflammation and secrete molecules involved in tissue repair. Overabundance of these repair molecules plays a role in certain disorders. In o autoimmune diseases such as rheumatoid arthritis, secretion of the cysteine peptidase cathepsin C degrades collagen, laminin, elastin and other structural proteins found in the extracellular matrix of bones.
  • Aspartic proteases are members of the cathepsin family of lysosomal proteases and include pepsin A, gastricsin, chymosin, renin, and cathepsins D and E. Aspartic proteases have a pair of 5 aspartic acid residues in the active site, and are most active in the pH 2 - 3 range, in which one of the aspartate residues is ionized, the other un-ionized. Aspartic proteases include bacterial penicillopepsin, mammalian pepsin, renin, chymosin, and certain fungal proteases. Abnormal regulation and expression of cathepsins is evident in various inflammatory disease states.
  • the mRNA for stromelysin, cytokines, TIMP-1, cathepsin, gelatinase, o and other molecules is preferentially expressed.
  • Expression of cathepsins L and D is elevated in synovial tissues from patients with rheumatoid arthritis and osteoarthritis.
  • Cathepsin L expression may also contribute to the influx of mononuclear cells which exacerbates the destruction of the rheumatoid synovium. (Keyszer, G.M. (1995) Arthritis Rheum.
  • Metalloproteases have active sites that include two glutamic acid residues and one histidine residue that serve as binding sites for zinc.
  • Carboxypeptidases A and B are the principal mammalian metalloproteases. Both are exoproteases of similar structure and active sites.
  • Carboxypeptidase A o like chymotrypsin, prefers C-terminal aromatic and aliphatic side chains of hydrophobic nature, whereas carboxypeptidase B is directed toward basic arginine and lysine residues.
  • Glycoprotease (GCP), or O-sialoglycoprotein endopeptidase is a metallopeptidase which specifically cleaves 0-sialoglycoproteins such as glycophorin A.
  • P-LAP placental leucine aminopeptidase
  • Ubiquitin proteases are associated with the ubiquitin conjugation system (UCS), a major pathway for the degradation of cellular proteins in eukaryotic cells and some bacteria.
  • the UCS mediates the elimination of abnormal proteins and regulates the half-lives of important regulatory proteins that control cellular processes such as gene transcription and cell cycle progression.
  • proteins targeted for degradation are conjugated to a ubiquitin, a small heat stable protein.
  • the ubiquitinated protein is then recognized and degraded by proteasome, a large, multisubunit proteolytic enzyme complex, and ubiquitin is released for reutilization by ubiquitin protease.
  • the UCS is implicated in the degradation of mitotic cyclic kinases, oncoproteins, tumor suppressor genes such as p53, viral proteins, cell surface receptors associated with signal transduction, transcriptional regulators, and mutated or damaged proteins (Ciechanover, A. (1994) Cell 79:13-21).
  • a murine proto-oncogene, Unp encodes a nuclear ubiquitin protease whose overexpression leads to oncogenic transformation of NIH3T3 cells, and the human homolog of this gene is consistently elevated in small cell tumors and adenocarcinomas of the lung (Gray, D. A. (1995) Oncogene 10:2179-2183).
  • the mechanism for the translocation process into the endoplasmic reticulum involves the recognition of an N-terminal signal peptide on the elongating protein.
  • the signal peptide directs the protein and attached ribosome to a receptor on the ER membrane.
  • the polypeptide chain passes through a pore in the ER membrane into the lumen while the N-terminal signal peptide remains attached at the membrane surface.
  • the process is completed when signal peptidase located inside the ER cleaves the signal peptide from the protein and releases the protein into the lumen.
  • Protease Inhibitors Protease inhibitors and other regulators of protease activity control the activity and effects of proteases.
  • Protease inhibitors have been shown to control pathogenesis in animal models of proteolytic disorders (Mu ⁇ hy, G. (1991) Agents Actions Suppl. 35:69-76). Low levels of the cystatins, low molecular weight inhibitors of the cysteine proteases, correlate with malignant progression of tumors. (Calkins, C. et al (1995) Biol. Biochem. Hoppe Seyler 376:71-80). Se ⁇ ins are inhibitors of mammalian plasma serine proteases. Many se ⁇ ins serve to regulate the blood clotting cascade and/or the complement cascade in mammals.
  • Sp32 is a positive regulator of the mammalian acrosomal protease, acrosin, that binds the proenzyme, proacrosin, and thereby aides in packaging the enzyme into the acrosomal matrix (Baba, T. et al. (1994) J. Biol. Chem. 269:10133- 10140).
  • the Kunitz family of serine protease inhibitors are characterized by one or more "Kunitz domains" containing a series of cysteine residues that are regularly spaced over approximately 50 amino acid residues and form three intrachain disulfide bonds.
  • TFPI-1 and TFPI-2 tissue factor pathway inhibitor
  • bikunin aprotinin
  • TFPI-1 and TFPI-2 tissue factor pathway inhibitor
  • inter- ⁇ -trypsin inhibitor inter- ⁇ -trypsin inhibitor
  • bikunin bikunin.
  • kalUkrein and plasmin serine proteases
  • Aprotinin 5 has cUnical utiUty in reduction of perioperative blood loss.
  • Protein folding in the ER is aided by two principal types of protein isomerases, protein disulfide isomerase (PDI), and peptidyl-prolyl isomerase (PPI)- PDI catalyzes the oxidation of free sulfhydryl groups in cysteine residues to form intramolecular disulfide bonds in proteins.
  • PDI protein disulfide isomerase
  • PPI peptidyl-prolyl isomerase
  • cyclophilins represent a major class of PPI that was originally identified as the major receptor for the immunosuppressive drug cyclosporin A (Handschumacher, RE. et al. (1984) Science 226: 544-547). 0 Protein Glvcosylation
  • O-linked glycosylation of proteins also occurs in the ER by the addition of N-acetylgalactosamine to the hydroxyl group of a serine or threonine residue followed by the sequential addition of other sugar o residues to the first. This process is catalysed by a series of glycosyltransferases each specific for a particular donor sugar nucleotide and acceptor molecule (Lodish, H. et al. (1995) Molecular Cell Biology, W.H. Freeman and Co., New York NY, pp.700-708). In many cases, both N- and O-linked oligosaccharides appear to be required for the secretion of proteins or the movement of plasma membrane glycoproteins to the cell surface.
  • An additional glycosylation mechanism operates in the ER specifically to target lysosomal enzymes to lysosomes and prevent their secretion.
  • Lysosomal enzymes in the ER receive an N-Unked oligosaccharide, like plasma membrane and secreted proteins, but are then phosphorylated on one or two mannose residues.
  • the phosphorylation of mannose residues occurs in two steps, the first step 5 being the addition of an N-acetylglucosamine phosphate residue by N-acetylglucosamine phosphotransferase, and the second the removal of the N-acetylglucosamine group by phosphodiesterase.
  • Chaperones o Molecular chaperones are proteins that aid in the proper folding of immature proteins and refolding of improperly folded ones, the assembly of protein subunits, and in the transport of unfolded proteins across membranes. Chaperones are also called heat-shock proteins (hsp) because of their tendency to be expressed in dramatically increased amounts following brief exposure of cells to elevated temperatures. This latter property most likely reflects their need in the refolding of proteins 5 that have become denatured by the high temperatures.
  • hsp heat-shock proteins
  • Chaperones may be divided into several classes according to their location, function, and molecular weight, and include hsp60, TCP1, hsp70, hsp40 (also called DnaJ), and hsp90.
  • hsp90 binds to steroid hormone receptors, represses transcription in the absence of the Ugand, and provides proper folding of the ligand-binding domain of the receptor in the presence of the hormone (Burston, S.G. and A.R. Clarke (1995) Essays 0 Biochem. 29:125-136).
  • Hsp60 and hsp70 chaperones aid in the transport and folding of newly synthesized proteins.
  • Hsp70 acts early in protein folding, binding a newly synthesized protein before it leaves the ribosome and transporting the protein to the mitochondria or ER before releasing the folded protein.
  • Hsp60 along with hsp 10, binds misfolded proteins and gives them the opportunity to refold correctly.
  • Al chaperones share an affinity for hydrophobic patches on incompletely folded 5 proteins and the ability to hydrolyze ATP. The energy of ATP hydrolysis is used to release the hsp- bound protein in its properly folded state (Aberts, supra, pp 214, 571-572).
  • SEQ ID NO:35 and SEQ ID NO:36 encode, for example, nucleic acid synthesis and o modification molecules .
  • DNA and RNA replication are critical processes for cell replication and function.
  • DNA and RNA replication are mediated by the enzymes DNA and RNA polymerase, respectively, by a "templating" process in which the nucleotide sequence of a DNA or RNA strand is copied by 5 complementary base-pairing into a complementary nucleic acid sequence of either DNA or RNA.
  • templating the process in which the nucleotide sequence of a DNA or RNA strand is copied by 5 complementary base-pairing into a complementary nucleic acid sequence of either DNA or RNA.
  • DNA polymerase catalyzes the stepwise addition of a deoxyribonucleotide to the 3' -OH end of a polynucleotide strand (the primer strand) that is paired to a second (template) strand.
  • the new DNA strand therefore grows in the 5' to 3' direction (Alberts, B. et al. (1994)The Molecular Biology 5 of the Cell, Garland Publishing Inc., New York NY, pp. 251-254).
  • the substrates for the polymerization reaction are the corresponding deoxynucleotide triphosphates which must base-pair with the correct nucleotide on the template strand in order to be recognized by the polymerase.
  • each of the two strands may serve as a template for the formation of a new complementary strand.
  • Each of the two daughter cells of the dividing cell 0 therefore inherits a new DNA double helix containing one old and one new strand.
  • DNA is said to be replicated "semiconservatively" by DNA polymerase.
  • DNA polymerase is also involved in the repair of damaged DNA as discussed below under “Ligases.”
  • RNA polymerase uses a DNA template strand to "transcribe" 5 DNA into RNA using ribonucleotide triphosphates as substrates. Like DNA polymerization, RNA polymerization proceeds in a 5' to 3' direction by addition of a ribonucleoside monophosphate to the 3' -OH end of a growing RNA chain. DNA transcription generates messenger RNAs (mRNA) that carry information for protein synthesis, as well as the transfer, ribosomal, and other RNAs that have structural or catalytic functions. In eukaryotes, three discrete RNA polymerases synthesize the three o different types of RNA (Aberts, supra, pp. 367-368).
  • mRNA messenger RNAs
  • RNA polymerase I makes the large ribosomal RNAs
  • RNA polymerase II makes the mRNAs that will be translated into proteins
  • RNA polymerase III makes a variety of small, stable RNAs, including 5S ribosomal RNA and the transfer RNAs (tRNA).
  • RNA synthesis is initiated by binding of the RNA polymerase to a promoter region on the DNA and synthesis begins at a start site within the promoter. Synthesis is 5 completed at a broad, general stop or termination region in the DNA where both the polymerase and the completed RNA chain are released.
  • DNA repair is the process by which accidental base changes, such as those produced by oxidative damage, hydrolytic attack, or uncontrolled methylation of DNA are corrected before o replication or transcription of the DNA can occur. Because of the efficiency of the DNA repair process, fewer than one in one thousand accidental base changes causes a mutation (Alberts, supra, pp. 245-249).
  • the three steps common to most types of DNA repair are (1) excision of the damaged or altered base or nucleotide by DNA nucleases, leaving a gap; (2) insertion of the correct nucleotide in this gap by DNA polymerase using the complementary strand as the template; and (3) sealing the 5 break left between the inserted nucleotide(s) and the existing DNA strand by DNA ligase.
  • DNA ligase uses the energy from ATP hydrolysis to activate the 5 ' end of the broken phosphodiester bond before forming the new bond with the 3'-OH of the DNA strand.
  • Bloom's syndrome an inherited human disease, individuals are partially deficient in DNA ligation and consequently have an increased incidence of cancer (Alberts, supra, p. 247). 5 Nucleases
  • Nucleases comprise both enzymes that hydrolyze DNA (DNase) and RNA (RNase). They serve different pu ⁇ oses in nucleic acid metaboUsm. Nucleases hydrolyze the phosphodiester bonds between adjacent nucleotides either at internal positions (endonucleases) or at the terminal 3' or 5' nucleotide positions (exonucleases).
  • a DNA exonuclease activity in DNA polymerase for example, o serves to remove improperly paired nucleotides attached to the 3'-OH end of the growing DNA strand by the polymerase and thereby serves a "proofreading" function. As mentioned above, DNA endonuclease activity is involved in the excision step of the DNA repair process.
  • RNases also serve a variety of functions.
  • RNase P is a ribonucleoprotein enzyme which cleaves the 5' end of pre-tRNAs as part of their maturation process.
  • RNase H digests 5 the RNA strand of an RN A/DNA hybrid. Such hybrids occur in cells invaded by retroviruses, and RNase H is an important enzyme in the retroviral replication cycle.
  • Pancreatic RNase secreted by the pancreas into the intestine hydrolyzes RNA present in ingested foods.
  • RNase activity in serum and cell extracts is elevated in a variety of cancers and infectious diseases (Schein, CH. (1997) Nat. Biotechnol. 15:529-536). Regulation of RNase activity is being investigated as a means to control o tumor angiogenesis, allergic reactions, viral infection and replication, and fungal infections.
  • Methylation of specific nucleotides occurs in both DNA and RNA, and serves different functions in the two macromolecules. Methylation of cytosine residues to form 5-methyl cytosine in DNA occurs specifically at CG sequences which are base-paired with one another in the DNA double- 5 helix. This pattern of methylation is passed from generation to generation during DNA replication by an enzyme called "maintenance methylase" that acts preferentially on those CG sequences that are base-paired with a CG sequence that is already methylated. Such methylation appears to distinguish active from inactive genes by preventing the binding of regulatory proteins that "turn on” the gene, but permit the binding of proteins that inactivate the gene (Alberts, supra, pp. 448-451).
  • tRNA methylase produces one of several nucleotide modifications in tRNA that affect the conformation and base-pairing of the molecule and facilitate the recognition of the appropriate mRNA codons by specific tRNAs.
  • the primary methylation pattern is the dimethylation of guanine residues to form N,N-dimethyl guanine.
  • HeUcases and Single-Stranded Binding Proteins 5 HeUcases are enzymes that destabiUze and unwind double helix structures in both DNA and RNA. Since DNA replication occurs more or less simultaneously on both strands, the two strands must first separate to generate a replication "fork” for DNA polymerase to act on.
  • DNA heUcases hydrolyze ATP and use the energy of hydrolysis to separate the DNA strands.
  • Single- 5 stranded binding proteins SSBs then bind to the exposed DNA strands without covering the bases, thereby temporarily stabilizing them for templating by the DNA polymerase (Alberts, supra, pp. 255- 256).
  • RNA helicases also alter and regulate RNA conformation and secondary structure. Like the DNA helicases, RNA helicases utilize energy derived from ATP hydrolysis to destabilize and unwind l o RNA duplexes.
  • the most well-characterized and ubiquitous family of RNA helicases is the DEAD- box family, so named for the conserved B-type ATP-binding motif which is diagnostic of proteins in this family.
  • DEAD-box helicases Over 40 DEAD-box helicases have been identified in organisms as diverse as bacteria, insects, yeast, amphibians, mammals, and plants. DEAD-box helicases function in diverse processes such as translation initiation, splicing, ribosome assembly, and RNA editing, transport, and stability.
  • DEAD-box helicases play tissue- and stage-specific roles in spermatogenesis and embryogenesis.
  • Overexpression of the DEAD-box 1 protein (DDX1) may play a role in the progression of neuroblastoma (Nb) and retinoblastoma (Rb) tumors (Godbout, R. et al. (1998) J. Biol. Chem. 273:21161-21168). These observations suggest that DDX1 may promote or enhance tumor progression by altering the normal secondary structure and expression levels of RNA in cancer cells.
  • DNA topoisomerase effectively acts as a reversible nuclease that hydrolyzes a phosphodiesterase bond in a DNA strand, permitting the two strands to
  • DNA topoisomerase I causes a single-strand break in a DNA helix to allow the rotation of the two strands of the helix about the remaining phosphodiester bond in the opposite strand.
  • DNA topoisomerase II causes a transient break in both strands of a DNA helix where two double heUces 35 cross over one another. This type of topoisomerase can efficiently separate two interlocked DNA circles (Aberts, supra, pp.260-262).
  • Topoisomerase II has been implicated in multi-drug resistance (MDR) as it appears to aid in the repair of DNA damage inflicted by DNA binding agents such as doxorubicin and vincristine. Recombinases
  • Genetic recombination is the process of rearranging DNA sequences within an organism's genome to provide genetic variation for the organism in response to changes in the environment.
  • DNA recombination allows variation in the particular combination of genes present in an individual's genome, as well as the timing and level of expression of these genes (see Aberts, supra, pp. 263- 273).
  • Two broad classes of genetic recombination are commonly recognized, general recombination and site-specific recombination.
  • General recombination involves genetic exchange between any homologous pair of DNA sequences usually located on two copies of the same chromosome.
  • recombinases that "nick" one strand of a DNA duplex more or less randomly and permit exchange with the complementary strand of another duplex.
  • the process does not normally change the arrangement of genes on a chromosome.
  • the recombinase recognizes specific nucleotide sequences present in one or both of the recombining molecules. Base-pairing is not involved in this form of recombination and therefore does not require DNA homology between the recombining molecules.
  • this form of recombination can alter the relative positions of nucleotide sequences in chromosomes.
  • RNA processing steps include capping at the 5' end with methylguanosine, polyadenylating the 3' end, and splicing to remove introns.
  • the primary RNA transcript from DNA is a faithful copy of the gene containing both exon and intron sequences, and the latter sequences must be cut out of the RNA transcript to produce an mRNA that codes for a protein.
  • This "splicing" of the mRNA sequence takes place in the nucleus with the aid of a large, multicomponent ribonucleoprotein complex known as a sphceosome.
  • the spUceosomal complex is composed of five small nuclear ribonucleoprotein particles (snRNPs) designated Ul, U2, U4, U5, and U6, and a number of additional proteins.
  • snRNP small nuclear ribonucleoprotein particles
  • Ul small nuclear ribonucleoprotein particles
  • U2, U4, U5, and U6 small nuclear ribonucleoprotein particles
  • U6 small nuclear ribonucleoprotein particles
  • RNA components of some snRNPs recognize and base pair with intron consensus sequences.
  • the protein components mediate sphceosome assembly and the splicing reaction.
  • Autoantibodies to snRNP proteins are found in the blood of patients with systemic lupus erythematosus (Stryer, L. (1995) Biochemistry, W.H. Freeman and Company, New York NY, p. 863).
  • Adhesion Molecules The surface of a cell is rich in transmembrane proteoglycans, glycoproteins, glycoUpids, and receptors. These macromolecules mediate adhesion with other cells and with components of the extracellular matrix (ECM). The interaction of the cell with its surroundings profoundly influences cell shape, strength, flexibility, motiUty, and adhesion. These dynamic properties are intimately 5 associated with signal transduction pathways controlUng cell proliferation and differentiation, tissue construction, and embryonic development. Cadherins
  • Cadherins comprise a family of calcium-dependent glycoproteins that function in mediating cell-cell adhesion in virtually all soUd tissues of multicellular organisms. These proteins share o multiple repeats of a cadherin-specific motif, and the repeats form the folding units of the cadherin extracellular domain. Cadherin molecules cooperate to form focal contacts, or adhesion plaques, between adjacent epithelial cells.
  • the cadherin family includes the classical cadherins and protocadherins.
  • Classical cadherins include the E-cadherin, N-cadherin, and P-cadherin subfamilies. E-cadherin is present on many types of epithelial cells and is especially important for embryonic 5 development.
  • N-cadherin is present on nerve, muscle, and lens cells and is also critical for embryonic development.
  • P-cadherin is present on cells of the placenta and epidermis. Recent studies report that protocadherins are involved in a variety of cell-cell interactions (Suzuki, S.T. (1996) J. Cell Sci. 109:2609-2611).
  • the intracellular anchorage of cadherins is regulated by their dynamic association with catenins, a family of cytoplasmic signal transduction proteins associated with the actin o cytoskeleton.
  • cadherins The anchorage of cadherins to the actin cytoskeleton appears to be regulated by protein tyrosine phosphorylation, and the cadherins are the target of phosphorylation-induced junctional disassembly (Aberle, H. et al. (1996) J. Cell. Biochem. 61:514-523).
  • Integrins are ubiquitous transmembrane adhesion molecules that link the ECM to the internal 5 cytoskeleton. Integrins are composed of two noncovalently associated transmembrane glycoprotein subunits called and ⁇ . Integrins function as receptors that play a role in signal transduction. For example, binding of integrin to its extracellular ligand may stimulate changes in intracellular calcium levels or protein kinase activity (Sjaastad, M.D. and W.J. Nelson (1997) BioEssays 19:47-55). At least ten cell surface receptors of the integrin family recognize the ECM component fibronectin, o which is involved in many different biological processes including cell migration and embryogenesis
  • Lectins comprise a ubiquitous family of extracellular glycoproteins which bind cell surface carbohydrates specifically and reversibly, resulting in the agglutination of cells (reviewed in 5 Drickamer, K. and M.E. Taylor (1993) Annu. Rev. Cell Biol. 9:237-264). This function is particularly important for activation of the immune response. Lectins mediate the agglutination and mitogenic stimulation of lymphocytes at sites of inflammation (Lasky, L.A. (1991) J. Cell. Biochem. 45:139-146; Paietta, E. et al. (1989) J. Immunol. 143:2850-2857).
  • Lectins are further classified into subfamiUes based on carbohydrate-binding specificity and 5 other criteria.
  • the galectin subfamily includes lectins that bind ⁇ -galactoside carbohydrate moieties in a thiol-dependent manner (reviewed in Hadari, Y.R. et al. (1998) J. Biol. Chem. 270:3447-3453). Galectins are widely expressed and developmentally regulated. Because all galectins lack an N-terminal signal peptide, it is suggested that galectins are externalized through an atypical secretory mechanism. Two classes of galectins have been defined based on molecular weight 0 and oUgomerization properties. Small galectins form homodimers and are about 14 to 16 kilodaltons in mass, while large galectins are monomeric and about 29-37 kilodaltons.
  • Galectins contain a characteristic carbohydrate recognition domain (CRD).
  • the CRD is about 140 amino acids and contains several stretches of about 1 - 10 amino acids which are highly conserved among all galectins.
  • a particular 6-amino acid motif within the CRD contains conserved 5 tryptophan and arginine residues which are critical for carbohydrate binding.
  • the CRD of some galectins also contains cysteine residues which may be important for disulfide bond formation. Secondary structure predictions indicate that the CRD forms several ⁇ -sheets.
  • Galectins play a number of roles in diseases and conditions associated with cell-cell and cell- matrix interactions. For example, certain galectins associate with sites of inflammation and bind to o cell surface immunoglobulin E molecules. In addition, galectins may play an important role in cancer metastasis. Galectin overexpression is correlated with the metastatic potential of cancers in humans and mice. Moreover, anti-galectin antibodies inhibit processes associated with cell transformation, such as cell aggregation and anchorage-independent growth (See, for example, Su, Z.-Z. et al. (1996) Proc. Natl. Acad. Sci. USA 93:7252-7257). 5 Selectins
  • Selectins comprise a specialized lectin subfamily involved primarily in inflammation and leukocyte adhesion (Reviewed in Lasky, supra). Selectins mediate the recruitment of leukocytes from the circulation to sites of acute inflammation and are expressed on the surface of vascular endothelial cells in response to cytokine signaling. Selectins bind to specific ligands on the o leukocyte cell membrane and enable the leukocyte to adhere to and migrate along the endothelial surface. Binding of selectin to its ligand leads to polarized rearrangement of the actin cytoskeleton and stimulates signal transduction within the leukocyte (Brenner, B. et al. (1997) Biochem. Biophys. Res.
  • selectins include lymphocyte adhesion molecule- 1 (Lam-1 or L-selectin), endothelial leukocyte adhesion molecule-1 (ELAM-1 or E-selectin), and granule membrane protein- 140 (GMP-140 or P-selectin) (Johnston, G.I. et al. (1989) Cell 56:1033-1044).
  • SEQ ID NO:37 encodes, for example, an antigen recognition molecule.
  • Al vertebrates have developed sophisticated and complex immune systems that provide protection from viral, bacterial, fungal, and parasitic infections.
  • a key feature of the immune system is its ability to distinguish foreign molecules, or antigens, from "self molecules. This ability is mediated primarily by secreted and transmembrane proteins expressed by leukocytes (white blood cells) such as lymphocytes, granulocytes, and monocytes. Most of these proteins belong to the immunoglobulin (Ig) superfamily, members of which contain one or more repeats of a conserved structural domain. This Ig domain is comprised of antiparallel ⁇ sheets joined by a disulfide bond in an arrangement called the Ig fold.
  • Members of the Ig superfamily include T-cell receptors, major histocompatibiUty (MHC) proteins, antibodies, and immune cell-specific surface markers such as CD4, CD8, and CD28.
  • MHC proteins are cell surface markers that bind to and present foreign antigens to T cells. MHC molecules are classified as either class I or class II. Class I MHC molecules (MHC I) are expressed on the surface of almost all cells and are involved in the presentation of antigen to cytotoxic T cells. For example, a cell infected with virus will degrade intracellular viral proteins and express the protein fragments bound to MHC I molecules on the cell surface. The MHC I/antigen complex is recognized by cytotoxic T-cells which destroy the infected cell and the virus within. Class II MHC molecules are expressed primarily on specialized antigen-presenting cells of the immune system, such as B-cells and macrophages.
  • MHC molecules also play an important role in organ rejection following transplantation. Rejection occurs when the recipient's T-cells respond to foreign MHC molecules on the transplanted organ in the same way as to self MHC molecules bound to foreign antigen. (Reviewed in Aberts, B. et al. (1994) Molecular Biology of the Cell. Garland PubUshing, New York NY, pp. 1229-1246.)
  • Aitibodies, or immunoglobulins are either expressed on the surface of B-cells or secreted by B-cells into the circulation. Aitibodies bind and neutralize foreign antigens in the blood and other extracellular fluids.
  • the prototypical antibody is a tetramer consisting of two identical heavy polypeptide chains (H-chains) and two identical light polypeptide chains (L-chains) interlinked by disulfide bonds. This arrangement confers the characteristic Y-shape to antibody molecules.
  • Antibodies are classified based on their H-chain composition.
  • the five antibody classes, IgA, IgD, IgE, IgG and IgM are defined by the ⁇ , ⁇ , e, ⁇ , and ⁇ H-chain types.
  • L- chains There are two types of L- chains, K and ⁇ , either of which may associate as a pair with any H-chain pair.
  • IgG the most 5 common class of antibody found in the circulation, is tetrameric, while the other classes of antibodies are generally variants or multimers of this basic structure.
  • H-chains and L-chains each contain an N-terminal variable region and a C-terminal constant region.
  • the constant region consists of about 110 amino acids in L-chains and about 330 or 440 amino acids in H-chains.
  • the amino acid sequence of the constant region is nearly identical among 0 H- or L-chains of a particular class.
  • the variable region consists of about 110 amino acids in both H- and L-chains. However, the amino acid sequence of the variable region differs among H- or L-chains of a particular class.
  • Within each H- or L-chain variable region are three hypervariable regions of extensive sequence diversity, each consisting of about 5 to 10 amino acids. In the antibody molecule, the H- and L-chain hypervariable regions come together to form the antigen recognition site. 5 (Reviewed in Aberts, supra, pp. 1206-1213 and 1216-1217.)
  • Both H-chains and L-chains contain repeated Ig domains.
  • a typical H-chain contains four Ig domains, three of which occur within the constant region and one of which occurs within the variable region and contributes to the formation of the antigen recognition site.
  • a typical L-chain contains two Ig domains, one of which occurs within the constant region and one of o which occurs within the variable region.
  • the immune system is capable of recognizing and responding to any foreign molecule that enters the body. Therefore, the immune system must be armed with a full repertoire of antibodies against all potential antigens.
  • antibody diversity is generated by somatic rearrangement of gene segments encoding variable and constant regions. These gene segments are joined together by site- 5 specific recombination which occurs between highly conserved DNA sequences that flank each gene segment. Because there are hundreds of different gene segments, millions of unique genes can be generated combinatorially. In addition, imprecise joining of these segments and an unusually high rate of somatic mutation within these segments further contribute to the generation of a diverse antibody population.
  • o T-cell receptors are both structurally and functionally related to antibodies. (Reviewed in
  • T-cell receptors are cell surface proteins that bind foreign antigens and mediate diverse aspects of the immune response.
  • a typical T-cell receptor is a heterodimer comprised of two disulfide-Unked polypeptide chains called ⁇ and ⁇ . Each chain is about 280 amino acids in length and contains one variable region and one constant region. Each variable or constant region folds 5 into an Ig domain. The variable regions from the ⁇ and ⁇ chains come together in the heterodimer to form the antigen recognition site.
  • T-cell receptor diversity is generated by somatic rearrangement of gene segments encoding the ⁇ and ⁇ chains.
  • T-cell receptors recognize small peptide antigens that are expressed on the surface of antigen-presenting cells and pathogen-infected cells. These peptide antigens are presented on the cell surface in association with major histocompatibiUty proteins which provide the 5 proper context for antigen recognition.
  • SEQ ID NO:38 and SEQ ID NO:39 encode, for example, secreted/extracellular matrix molecules.
  • Protein secretion is essential for cellular function. Protein secretion is mediated by a signal peptide located at the amino terminus of the protein to be secreted.
  • the signal peptide is comprised of about ten to twenty hydrophobic amino acids which target the nascent protein from the ribosome to the endoplasmic reticulum (ER). Proteins targeted to the ER may either proceed through the secretory pathway or remain in any of the secretory organelles such as theER, Golgi apparatus, or lysosomes. 5 Proteins that transit through the secretory pathway are either secreted into the extracellular space or retained in the plasma membrane.
  • Secreted proteins are often synthesized as inactive precursors that are activated by post-translational processing events during transit through the secretory pathway. Such events include glycosylation, proteolysis, and removal of the signal peptide by a signal peptidase. Other events that may occur during protein transport include chaperone-dependent unfolding and o folding of the nascent protein and interaction of the protein with a receptor or pore complex. Examples of secreted proteins with amino terminal signal peptides include receptors, extracellular matrix molecules, cytokines, hormones, growth and differentiation factors, neuropeptides, vasomediators, ion channels, transporters/pumps, and proteases. (Reviewed in Aberts, B. et al.
  • the extracellular matrix is a complex network of glycoproteins, polysaccharides, proteoglycans, and other macromolecules that are secreted from the cell into the extracellular space.
  • the ECM remains in close association with the cell surface and provides a supportive meshwork that profoundly influences cell shape, motility, strength, flexibility, and adhesion. In fact, adhesion of a cell to its surrounding matrix is required for cell survival except in the case of metastatic tumor cells, o which have overcome the need for cell-ECM anchorage.
  • the collagens comprise a family of ECM proteins that provide structure to bone, teeth, skin, Ugaments, tendons, cartilage, blood vessels, and basement membranes. Multiple collagen proteins have been identified. Three collagen molecules fold together in a triple helix stabihzed by interchain disulfide bonds. Bundles of these triple heUces then associate to form fibrils.
  • Collagen primary structure 5 consists of hundreds of (Gly-X-Y) repeats where about a third of the X and Y residues are Pro.
  • Glycines are crucial to helix formation as the bulkier amino acid sidechains cannot fold into the triple helical conformation. Because of these strict sequence requirements, mutations in collagen genes have severe consequences. Osteogenesis imperfecta patients have brittle bones that fracture easily; in severe cases patients die in utero or at birth. Ehlers-Danlos syndrome patients have hyperelastic skin, 0 hypermobile joints, and susceptibiUty to aortic and intestinal rupture. Chondrodysplasia patients have short stature and ocular disorders. Aport syndrome patients have hematuria, sensorineural deafness, and eye lens deformation. (Isselbacher, KJ. et al.
  • Elastin is a highly hydrophobic protein of about 750 amino acids that is rich in proUne and glycine residues. Elastin molecules are highly cross-Unked, forming an extensive extracellular network of fibers and sheets. Elastin fibers are surrounded by a sheath of microfibrils which are composed of a number of glycoproteins, including fibrilUn. Mutations in the gene encoding fibrilUn are responsible for o Marfan' s syndrome, a genetic disorder characterized by defects in connective tissue. In severe cases, the aortas of afflicted individuals are prone to rupture. (Reviewed in Aberts, supra, pp. 984-986.)
  • Fibronectin is a large ECM glycoprotein found in all vertebrates. Fibronectin exists as a dimer of two subunits, each containing about 2,500 amino acids. Each subunit folds into a rod-Uke structure containing multiple domains. The domains each contain multiple repeated modules, the most common 5 of which is the type III fibronectin repeat.
  • the type III fibronectin repeat is about 90 amino acids in length and is also found in other ECM proteins and in some plasma membrane and cytoplasmic proteins.
  • some type III fibronectin repeats contain a characteristic tripeptide consisting of Aginine-Glycine-Aspartic acid (RGD). The RGD sequence is recognized by the integrin family of cell surface receptors and is also found in other ECM proteins.
  • Laminin is a major glycoprotein component of the basal lamina which underlies and supports epitheUal cell sheets.
  • Laminin is one of the first ECM proteins synthesized in the developing embryo.
  • Laminin is an 850 kilodalton protein composed of three polypeptide chains joined in the shape of a cross by disulfide bonds.
  • Laminin is especially important for angiogenesis and in particular, for guiding the formation of capillaries. (Reviewed in Aberts, supra, pp. 990-991.)
  • proteoglycans are composed of unbranched polysaccharide chains
  • proteoglycans attached to protein cores.
  • Common proteoglycans include aggrecan, betaglycan, decorin, perlecan, serglycin, and syndecan-1.
  • Some of these molecules not only provide mechanical support, but also bind to extracellular signaUng molecules, such as fibroblast growth factor and transforming growth factor ⁇ , suggesting a role for proteoglycans in cell-cell communication and cell 0 growth.
  • glycoproteins tenascin-C and tenascin-R are expressed in developing and lesioned neural tissue and provide stimulatory and anti- adhesive (inhibitory) properties, respectively, for axonal growth. (Faissner, A. (1997) Cell Tissue Res. 290:331-341.)
  • SEQ ID NO:40, SEQ ID NO:41, SEQ ID NO:42, SEQ ID NO:43, SEQ ID NO:44, and SEQ ID NO:45 encode, for example, cytoskeletal molecules.
  • the cytoskeleton is a cytoplasmic network of protein fibers that mediate cell shape, structure, and movement.
  • the cytoskeleton supports the cell membrane and forms tracks along which o organelles and other elements move in the cytosol.
  • the cytoskeleton is a dynamic structure that allows cells to adopt various shapes and to carry out directed movements.
  • Major cytoskeletal fibers include the microtubules, the microfilaments, and the intermediate filaments.
  • Motor proteins including myosin, dynein, and kinesin, drive movement of or along the fibers.
  • the motor protein dynamin drives the formation of membrane vesicles. Accessory or associated proteins modify the 5 structure or activity of the fibers while cytoskeletal membrane anchors connect the fibers to the cell membrane.
  • Tubulins include myosin, dynein, and kinesin.
  • Microtubules cytoskeletal fibers with a diameter of about 24 nm, have multiple roles in the cell. Bundles of microtubules form cilia and flagella, which are whip-like extensions of the cell o membrane that are necessary for sweeping materials across an epithelium and for swimming of sperm, respectively. Marginal bands of microtubules in red blood cells and platelets are important for these cells' pliability. Organelles, membrane vesicles, and proteins are transported in the cell along tracks of microtubules. For example, microtubules run through nerve cell axons, allowing bidirectional transport of materials and membrane vesicles between the cell body and the nerve terminal. Failure to supply the nerve terminal with these vesicles blocks the transmission of neural signals. Microtubules are also critical to chromosomal movement during cell division. Both stable and short-lived populations of microtubules exist in the cell.
  • Microtubules are polymers of GTP-binding tubulin protein subunits. Each subunit is a 5 heterodimer of ⁇ - and ⁇ - tubulin, multiple isoforms of which exist.
  • the hydrolysis of GTP is linked to the addition of tubuUn subunits at the end of a microtubule.
  • the subunits interact head to tail to form protofilaments; the protofilaments interact side to side to form a microtubule.
  • a microtubule is polarized, one end ringed with ⁇ -tubulin and the other with ⁇ -tubulin, and the two ends differ in their rates of assembly.
  • each microtubule is composed of 13 protofilaments although 11 or 15 0 protofilament-microtubules are sometimes found.
  • Cilia and flagella contain doublet microtubules.
  • Microtubules grow from speciaUzed structures known as centrosomes or microtubule-organizing centers (MTOCs). MTOCs may contain one or two centrioles, which are pinwheel arrays of triplet microtubules.
  • the basal body, the organizing center located at the base of a cilium or flagellum, contains one centriole.
  • Gamma tubulin present in the MTOC is important for nucleating the 5 polymerization of ⁇ - and ⁇ - tubulin heterodimers but does not polymerize into microtubules.
  • Microtubule- Associated Proteins are important for nucleating the 5 polymerization of ⁇ - and ⁇ - tubulin heterodimers but does not polymerize into microtubules.
  • Microtubule-associated proteins have roles in the assembly and stabilization of microtubules.
  • assembly MAPs can be identified in neurons as well as non-neuronal cells. Assembly MAPs are responsible for cross-Unking microtubules in the cytosol. o These MAPs are organized into two domains: a basic microtubule-binding domain and an acidic projection domain. The projection domain is the binding site for membranes, intermediate filaments, or other microtubules. Based on sequence analysis, assembly MAPs can be further grouped into two types: Type I and Type II.
  • Type I MAPs which include MAPI A and MAP1B, are large, filamentous molecules that co-purify with microtubules and are abundantly expressed in brain and testes.
  • Type I 5 MAPs contain several repeats of a positively-charged amino acid sequence motif that binds and neutralizes negatively charged tubulin, leading to stabiUzation of microtubules.
  • MAPI A and MAP1B are each derived from a single precursor polypeptide that is subsequently proteolytically processed to generate one heavy chain and one Ught chain.
  • LC3 Another Ught chain, LC 3, is a 16.4 kDa molecule that binds MAPI A, MAP1B, and o microtubules. It is suggested that LC3 is synthesized from a source other than the MAPI A or MAPI B transcripts, and that the expression of LC3 may be important in regulating the microtubule binding activity of MAPI A and MAP1B during cell proUferation (Mann, S.S. et al. (1994) J. Biol. Chem. 269:11492-11497).
  • Type II MAPs which include MAP2a, MAP2b, MAP2c, MAP4, and Tau, are characterized by three to four copies of an 18-residue sequence in the microtubule-binding domain.
  • MAP2a, MAP2b, and MAP2c are found only in dendrites
  • MAP4 is found in non-neuronal cells
  • Tau is found in axons and dendrites of nerve cells. Ater native spUcing of the Tau mRNA leads to the existence of multiple forms of Tau protein.
  • Tau phosphorylation is altered in neurodegenerative disorders such as Azheimer' s disease, Pick's disease, progressive supranuclear palsy, corticobasal degeneration, and familial frontotemporal dementia and Parkinsonism linked to chromosome 17.
  • the altered Tau phosphorylation leads to a collapse of the microtubule network and the formation of intraneuronal Tau aggregates (Spillantini, M.G. and M. Goedert (1998) Trends Neurosci. 21:428-433).
  • the protein pericentrin is found in the MTOC and has a role in microtubule assembly. Actins
  • Microfilaments are vital to cell locomotion, cell shape, cell adhesion, cell division, and muscle contraction. Assembly and disassembly of the microfilaments allow cells to change their mo ⁇ hology. Microfilaments are the polymerized form of actin, the most abundant intracellular protein in the eukaryotic cell. Human cells contain six isoforms of actin. The three ⁇ -actins are found in different kinds of muscle, nonmuscle ⁇ - actin and nonmuscle ⁇ -actin are found in nonmuscle cells, and another ⁇ -actin is found in intestinal smooth muscle cells.
  • G-actin the monomeric form of actin, polymerizes into polarized, helical F- actin filaments, accompanied by the hydrolysis of ATP to ADP.
  • Actin filaments associate to form bundles and networks, providing a framework to support the plasma membrane and determine cell shape. These bundles and networks are connected to the cell membrane.
  • thin filaments containing actin slide past thick filaments containing the motor protein myosin during contraction.
  • a family of actin-related proteins exist that are not part of the actin cytoskeleton, but rather associate with microtubules and dynein.
  • Actin-associated proteins have roles in cross-linking, severing, and stabilization of actin filaments and in sequestering actin monomers.
  • actin-associated proteins have multiple functions. Bundles and networks of actin filaments are held together by actin cross-linking proteins. These proteins have two actin-binding sites, one for each filament. Short cross-linking proteins promote bundle formation while longer, more flexible cross-linking proteins promote network formation. Calmodulin-like calcium-binding domains in actin cross-linking proteins allow calcium regulation of cross-linking.
  • Group I cross-linking proteins have unique actin-binding domains and include the 30 kD protein, EF-la, fascin, and scruin.
  • Group II cross-linking proteins have a 7,000- MW actin-binding domain and include villin and dematin.
  • Group III cross-Unking proteins have pairs of a 26,000-MW actin-binding domain and include fimbrin, spectrin, dystrophin, ABP 120, and filamin.
  • Severing proteins regulate the length of actin filaments by breaking them into short pieces or by blocking their ends. Severing proteins include gCAP39, severin (fragmin), gelsolin, and vilUn.
  • Capping proteins can cap the ends of actin filaments, but cannot break filaments. Capping proteins include CapZ and tropomodulin.
  • Intermediate filaments are cytoskeletal fibers with a diameter of about 10 nm, intermediate between that of microfilaments and microtubules. IFs serve structural roles in the cell, reinforcing cells and organizing cells into tissues. IFs are particularly abundant in epidermal cells and in neurons. IFs are extremely stable, and, in contrast to microfilaments and microtubules, do not function in cell motility.
  • Type I and Type II proteins are the acidic and basic keratins, respectively.
  • Heterodimer s of the acidic and basic keratins are the building blocks of keratin IFs. Keratins are abundant in soft epitheUa such as skin and cornea, hard epitheUa such as nails and hair, and in epitheUa that Une internal body cavities.
  • Type III IF proteins include desmin, gUal fibrillary acidic protein, vimentin, and peripherin.
  • Desmin filaments in muscle cells link myofibrils into bundles and stabiUze sarcomeres in contracting muscle.
  • GUal fibrillary acidic protein filaments are found in the gUal cells that surround neurons and astrocytes.
  • Vimentin filaments are found in blood vessel endothelial cells, some epitheUal cells, and mesenchymal cells such as fibroblasts, and are commonly associated with microtubules. Vimentin filaments may have roles in keeping the nucleus and other organelles in place in the cell.
  • Type IV IFs include the neurofilaments and nestin.
  • Neurofilaments composed of three polypeptides NF-L, NF-M, and NF-H, are frequently associated with microtubules in axons. Neurofilaments are responsible for the radial growth and diameter of an axon, and ultimately for the speed of nerve impulse transmission. Changes in phosphorylation and metabolism of neurofilaments are observed in neurodegenerative diseases including amyotrophic lateral sclerosis, Parkinson's disease, and Azheimer 's disease (Julien, J.P. and W.E. Mushynski (1998) Prog. Nucleic Acid Res. Mol. Biol. 61:1-23). Type V IFs, the lamins, are found in the nucleus where they support the nuclear membrane.
  • IFs have a central ⁇ -heUcal rod region interrupted by short nonheUcal Unker segments.
  • the rod region is bracketed, in most cases, by non-heUcal head and tail domains.
  • the rod regions of intermediate filament proteins associate to form a coiled-coil dimer.
  • a highly ordered assembly process leads from the dimers to the IFs. Neither ATP nor GTP is needed for IF assembly, unUke that of 5 microfilaments and microtubules.
  • IF-associated proteins mediate the interactions of IFs with one another and with other cell structures.
  • IFAPs cross-link IFs into a bundle, into a network, or to the plasma membrane, and may cross-Unk IFs to the microfilament and microtubule cytoskeleton. Microtubules and IFs are in particular closely associated.
  • IFAPs include BPAG1, plakoglobin, desmoplakin I, desmoplakin II, o plectin, ankyrin, filaggrin, and lamin B receptor.
  • Cytoskeletal fibers are attached to the plasma membrane by specific proteins. These attachments are important for maintaining cell shape and for muscle contraction.
  • the spectrin- actin cytoskeleton is attached to cell membrane by three proteins, band 4.1, ankyrin, and 5 adducin. Defects in this attachment result in abnormally shaped cells which are more rapidly degraded by the spleen, leading to anemia.
  • the spectrin-actin cytoskeleton is also linked to the membrane by ankyrin; a second actin network is anchored to the membrane by filamin.
  • the protein dystrophin links actin filaments to the plasma membrane; mutations in the dystrophin gene lead to Duchenne muscular dystrophy.
  • adherens junctions and adhesion plaques o the peripheral membrane proteins ⁇ -actinin and vinculin attach actin filaments to the cell membrane.
  • IFs are also attached to membranes by cytoskeletal-membrane anchors.
  • the nuclear lamina is attached to the inner surface of the nuclear membrane by the lamin B receptor.
  • Vimentin IFs are attached to the plasma membrane by ankyrin and plectin.
  • Desmosome and hemidesmosome membrane junctions hold together epithelial cells of organs and skin. These membrane junctions 5 allow shear forces to be distributed across the entire epithelial cell layer, thus providing strength and rigidity to the epithelium.
  • IFs in epithelial cells are attached to the desmosome by plakoglobin and desmoplakins. The proteins that Unk IFs to hemidesmosomes are not known.
  • Desmin IFs surround the sarcomere in muscle and are linked to the plasma membrane by paranemin, synemin, and ankyrin.
  • Myosin-related Motor Proteins o Myosins are actin-activated ATPases, found in eukaryotic cells, that couple hydrolysis of
  • Myosin provides the motor function for muscle contraction and intracellular movements such as phagocytosis and rearrangement of cell contents during mitotic cell division (cytokinesis).
  • the contractile unit of skeletal muscle termed the sarcomere, consists of highly ordered arrays of thin actin-containing filaments and thick myosin-containing filaments. Crossbridges form between the thick and thin filaments, and the ATP-dependent movement of myosin heads within the thick filaments pulls the thin filaments, shortening the sarcomere and thus the muscle fiber.
  • Myosins are composed of one or two heavy chains and associated light chains.
  • Myosin heavy chains contain an amino-terminal motor or head domain, a neck that is the site of light-chain 5 binding, and a carboxy-terminal tail domain.
  • the tail domains may associate to form an ⁇ -helical coiled coil.
  • Conventional myosins such as those found in muscle tissue, are composed of two myosin heavy-chain subunits, each associated with two light-chain subunits that bind at the neck region and play a regulatory role.
  • Unconventional myosins believed to function in intracellular motion, may contain either one or two heavy chains and associated light chains. There is evidence for o about 25 myosin heavy chain genes in vertebrates, more than half of them unconventional.
  • Dyneins are (-) end-directed motor proteins which act on microtubules. Two classes of dyneins, cytosoUc and axonemal, have been identified. CytosoUc dyneins are responsible for translocation of materials along cytoplasmic microtubules, for example, transport from the nerve 5 terminal to the cell body and transport of endocytic vesicles to lysosomes. Cytoplasmic dyneins are also reported to play a role in mitosis. Axonemal dyneins are responsible for the beating of flagella and ciUa. Dynein on one microtubule doublet walks along the adjacent microtubule doublet.
  • Dyneins have a native mass between 1000 and 2000 kDa and contain either two or three force-producing heads driven by the o hydrolysis of ATP. The heads are Unked via stalks to a basal domain which is composed of a highly variable number of accessory intermediate and Ught chains.
  • Kinesins are (+) end-directed motor proteins which act on microtubules.
  • the prototypical kinesin molecule is involved in the transport of membrane-bound vesicles and organelles. This function 5 is particularly important for axonal transport in neurons.
  • Kinesin is also important in all cell types for the transport of vesicles from the Golgi complex to the endoplasmic reticulum. This role is critical for maintaining the identity and functionaUty of these secretory organelles.
  • Kinesins define a ubiquitous, conserved family of over 50 proteins that can be classified into at least 8 subfamiUes based on primary amino acid sequence, domain structure, velocity of movement, and 0 cellular function. (Reviewed in Moore, J.D. and S.A. Endow (1996) Bioessays 18:207-219; and Hoyt,
  • the prototypical kinesin molecule is a heterotetramer comprised of two heavy polypeptide chains (KHCs) and two light polypeptide chains (KLCs).
  • KHCs heavy polypeptide chains
  • KLCs light polypeptide chains
  • KHC subunits are typically referred to as "kinesin.” KHC is about 1000 amino acids in length, and
  • KLC is about 550 amino acids in length.
  • Two KHCs dimerize to form a rod-shaped molecule with three distinct regions of secondary structure.
  • At one end of the molecule is a globular motor domain that functions in ATP hydrolysis and microtubule binding.
  • Kinesin motor domains are highly conserved and share over 70% identity. Beyond the motor domain is an ⁇ -hehcal coiled-coil region which mediates dimerization.
  • a fan-shaped tail that associates with 5 molecular cargo. The tail is formed by the interaction of the KHC C-termini with the two KLCs.
  • KRPs kinesin-related proteins
  • Dynamin is a large GTPase motor protein that functions as a "molecular pinchase,” generating a mechanochemical force used to sever membranes. This activity is important in forming clathrin-coated vesicles from coated pits in endocytosis and in the biogenesis of synaptic vesicles in neurons.
  • dynamin' s self-assembly into spirals that may act to constrict a flat membrane surface into a tubule.
  • GTP hydrolysis induces a change in o conformation of the dynamin polymer that pinches the membrane tubule, leading to severing of the membrane tubule and formation of a membrane vesicle.
  • Release of GDP and inorganic phosphate leads to dynamin disassembly. Following disassembly the dynamin may either dissociate from the membrane or remain associated to the vesicle and be transported to another region of the cell.
  • Three homologous dynamin genes have been discovered, in addition to several dynamin-related proteins.
  • dynamin regions are the N-terminal GTP-binding domain, a central pleckstrin homology domain that binds membranes, a central coiled-coil region that may activate dynamin' s GTPase activity, and a C-terminal proline-rich domain that contains several motifs that bind SH3 domains on other proteins. Some dynamin-related proteins do not contain the pleckstrin homology domain or the proline-rich domain. (See McNiven, M.A. (1998) Cell 94:151-154; Scaife, R.M. and R.L. Margolis 0 (1997) Cell. Signal. 9:395-401.)
  • the cytoskeleton is reviewed in Lodish, H. et al. (1995) Molecular Cell Biology, Scientific American Books, New York NY.
  • Ribosomal Molecules SEQ ID NO:49, SEQ ID NO:50, SEQ ID NO:51, SEQ ID NO:52, and SEQ ID NO:53 encode, for example, ribosomal molecules.
  • Ribosomal RNAs are assembled, along with ribosomal proteins, into ribosomes, which are cytoplasmic particles that translate messenger RNA into polypeptides.
  • the eukaryotic ribosome is composed of a 60S (large) subunit and a 40S (small) subunit, which together form the 80S ribosome.
  • the ribosome also contains more than fifty proteins.
  • the ribosomal proteins have a prefix which denotes the subunit to which they belong, either L (large) or S (small).
  • Ribosomal protein activities include binding rRNA and organizing the conformation of the junctions between rRNA helices (Woodson, S.A. and N.B. Leontis (1998) Curr. Opin. Struct. Biol. 8:294-300; Ramakrishnan, V. and S.W. White (1998) Trends Biochem. Sci. 23:208-212.)
  • Three important sites are identified on the ribosome.
  • the aminoacyl- tRNA site (A site) is where charged tRNAs (with the exception of the initiator-tRNA) bind on arrival at the ribosome.
  • the peptidyl-tRNA site (P site) is where new peptide bonds are formed, as well as where the initiator tRNA binds.
  • the exit site is where deacylated tRNAs bind prior to their release from the ribosome.
  • the ribosome is reviewed in Stryer, L. (1995) Biochemistry W.H. Freeman and Company, New York NY, pp. 888-908; and Lodish, H. et al. (1995) Molecular Cell Biology Scientific American Books, New York NY. pp. 119-138.)
  • Chromatin Molecules The nuclear DNA of eukaryotes is organized into chromatin. Two types of chromatin are observed: euchromatin, some of which may be transcribed, and heterochromatin so densely packed that much of it is inaccessible to transcription. Chromatin packing thus serves to regulate protein expression in eukaryotes. Bacteria lack chromatin and the chromatin-packing level of gene regulation.
  • the fundamental unit of chromatin is the nucleosome of 200 DNA base pairs associated with two copies each of histones H2A, H2B, H3, and H4. Adjascent nucleosomes are Unked by another class of histones, HI .
  • HMG high mobiUty group
  • Chromodomain proteins function in compaction of chromatin into its transcriptionally silent heterochromatin form. During mitosis, all DNA is compacted into heterochromatin and transcription ceases.
  • Patterns of chromatin structure can be stably inherited, producing heritable patterns of gene expression.
  • one of the two X chromosomes in each female cell is inactivated by 5 condensation to heterochromatin during zygote development.
  • the inactive state of this chromosome is inherited, so that adult females are mosaics of clusters of paternal-X and maternal-X clonal cell groups.
  • the condensed X chromosome is reactivated in meiosis.
  • Chromatin is associated with disorders of protein expression such as thalassemia, a genetic anemia resulting from the removal of the locus control region (LCR) required for decondensation of the o globin gene locus.
  • LCR locus control region
  • Glucose is initially converted to pyruvate in the cytoplasm.
  • Fatty acids and pyruvate are transported to the mitochondria for complete oxidation to C0 2 coupled by enzymes to the transport of electrons from NADH and FADH 2 5 to oxygen and to the synthesis of ATP (oxidative phosphorylation) from ADP and P,.
  • Pyruvate is transported into the mitochondria and converted to acetyl-CoA for oxidation via the citric acid cycle, involving pyruvate dehydrogenase components, dihydrolipoyl fransacetylase, and dihydrolipoyl dehydrogenase.
  • Enzymes involved in the citric acid cycle include: citrate synthetase, aconitases, isocitrate dehydrogenase, alpha-ketoglutarate dehydrogenase complex including o transsuccinylases, succinyl CoA synthetase, succinate dehydrogenase, fumarases, and malate dehydrogenase.
  • Acetyl CoA is oxidized to C0 2 with concomitant formation of NADH, FADH 2 , and GTP.
  • oxidative phosphorylation the transfer of electrons from NADH and FADH 2 to oxygen by dehydrogenases is coupled to the synthesis of ATP from ADP and P, by the F Q F, ATPase complex in the mitochondrial inner membrane.
  • Enzyme complexes responsible for electron transport and ATP synthesis include the F Q F J ATPase complex, ubiquinone(CoQ)-cytochrome c reductase, ubiquinone reductase, cytochrome b, cytochrome c 1; FeS protein, and cytochrome c oxidase.
  • ATP synthesis requires membrane transport enzymes including the phosphate transporter and the ATP- ADP antiport protein.
  • the ATP-binding casette (ABC) superfamily has also been suggested 5 as belonging to the mitochondrial transport group (Hogue, D.L. et al. (1999) J. Mol. Biol. 285:379- 389). Brown fat uncoupling protein dissipates oxidative energy as heat, and may be involved the fever response to infection and trauma (Cannon, B. et al. (1998) Ann. NY Acad. Sci. 856:171-187).
  • Mitochondria are oval-shaped organelles comprising an outer membrane, a tightly folded inner membrane, an intermembrane space between the outer and inner membranes, and a matrix o inside the inner membrane.
  • the outer membrane contains many porin molecules that allow ions and charged molecules to enter the intermembrane space, while the inner membrane contains a variety of transport proteins that transfer only selected molecules.
  • Mitochondria are the primary sites of energy production in cells.
  • Mitochondria contain a small amount of DNA.
  • Human mitochondrial DNA encodes 13 5 proteins, 22 tRNAs, and 2 rRNAs.
  • Mitochondrial-DNA encoded proteins include NADH-Q reductase, a cytochrome reductase subunit, cytochrome oxidase subunits, and ATP synthase subunits.
  • Cytochrome b5 is a central electron donor for various reductive reactions occurring on the cytoplasmic surface of liver o endoplasmic reticulum. Cytochrome b5 has been found in Golgi, plasma, endoplasmic reticulum
  • Import of these preproteins from the cytoplasm requires a multisubunit protein complex in the outer membrane known as the translocase of outer mitochondrial membrane (TOM; previously o designated MOM; Pfanner, N. et al. (1996) Trends Biochem. Sci. 21 :51-52) and at least three inner membrane proteins which comprise the translocase of inner mitochondrial membrane (TIM; previously designated MIM; Pfanner, supra).
  • TOM translocase of outer mitochondrial membrane
  • TIM previously designated MIM; Pfanner, supra
  • An inside-negative membrane potential across the inner mitochondrial membrane is also required for preprotein import.
  • Preproteins are recognized by surface receptor components of the TOM complex and are translocated through a proteinaceous pore formed 5 by other TOM components. Proteins targeted to the matrix are then recognized by the import machinery of the TIM complex.
  • the import systems of the outer and inner membranes can function independently (Segui-Real, B. et al. (1993) EMBO J. 12:2211-22
  • leader peptide is cleaved by a signal peptidase to generate the mature protein.
  • Most leader peptides are removed in a one step process by a protease termed mitochondrial processing peptidase (MPP) (Paces, V. et al. (1993) Proc. Natl. Acad. Sci. USA 90:5355-5358).
  • MPP mitochondrial processing peptidase
  • a two-step process occurs in which MPP generates an intermediate precursor form which is cleaved by a second enzyme, mitochondrial intermediate peptidase, to generate the mature protein.
  • Mitochondrial dysfunction leads to impaired calcium buffering, generation of free radicals that may participate in deleterious intracellular and extracellular processes, changes in mitochondrial permeability and oxidative damage which is observed in several neurodegenerative diseases.
  • Neurodegenerative diseases Unked to mitochondrial dysfunction include some forms of Alzheimer's disease, Friedreich's ataxia, familial amyotrophic lateral sclerosis, and Huntington's disease (Beal, M.F. (1998) Biochim. Biophys. Acta 1366:211-213). The myocardium is heavily dependent on oxidative metabolism, so mitochondrial dysfunction often leads to heart disease (DiMauro, S. and M. Hirano (1998) Curr. Opin. Cardiol 13:190-197).
  • Mitochondria are implicated in disorders of cell proliferation, since they play an important role in a cell's decision to proliferate or self-destruct through apoptosis.
  • the oncoprotein Bcl-2 promotes cell proliferation by stabilizing mitochondrial membranes so that apoptosis signals are not released (Susin, S.A. (1998) Biochim. Biophys. Acta 1366:151-165).
  • SEQ ID NO:19, SEQ ID NO:20, SEQ ID NO:21, SEQ ID NO:22, SEQ ID NO:23, SEQ ID NO:24, SEQ ID NO:25, SEQ ID NO:26, SEQ ID NO:27, SEQ ID NO:28, SEQ ID NO:29, SEQ ID NO:30, SEQ ID NO:31, SEQ ID NO:32, and SEQ ID NO:33 encode, for example, transcription factor molecules.
  • Multicellular organisms are comprised of diverse cell types that differ dramatically both in structure and function.
  • the identity of a cell is determined by its characteristic pattern of gene expression, and different cell types express overlapping but distinctive sets of genes throughout development. Spatial and temporal regulation of gene expression is critical for the control of cell proliferation, cell differentiation, apoptosis, and other processes that contribute to organismal development.
  • gene expression is regulated in response to extracellular signals that mediate cell-cell communication and coordinate the activities of different cell types. Appropriate gene regulation also ensures that cells function efficiently by expressing only those genes whose functions are required at a given time.
  • Transcriptional regulatory proteins are essential for the control of gene expression. Some of these proteins function as transcription factors that initiate, activate, repress, or terminate gene transcription.
  • Transcription factors generally bind to the promoter, enhancer, and upstream regulatory regions of a gene in a sequence-specific manner, although some factors bind regulatory elements within or downstream of a gene's coding region. Transcription factors may bind to a specific region of DNA singly or as a complex with other accessory factors. (Reviewed in Lewin, B. (1990) Genes IV, Oxford University Press, New York NY, and Cell Press, Cambridge MA, pp. 554-570.)
  • the double helix structure and repeated sequences of DNA create topological and chemical features which can be recognized by transcription factors. These features are hydrogen bond donor and acceptor groups, hydrophobic patches, major and minor grooves, and regular, repeated stretches of sequence which induce distinct bends in the helix.
  • transcription factors recognize specific DNA sequence motifs of about 20 nucleotides in length. Multiple, adjacent transcription factor-binding motifs may be required for gene regulation.
  • DNA-binding structural motifs which comprise either ⁇ heUces or ⁇ sheets that bind to the major groove of DNA.
  • structural motifs are helix-turn-helix, zinc finger, leucine zipper, and helix-loop-helix. Proteins containing these motifs may act alone as monomers, or they may form homo- or heterodimers that interact with DNA.
  • the helix-turn-helix motif consists of two ⁇ helices connected at a fixed angle by a short chain of amino acids. One of the helices binds to the major groove. Helix-turn-helix motifs are exemplified by the homeobox motif which is present in homeodomain proteins. These proteins are critical for specifying the anterior-posterior body axis during development and are conserved throughout the animal kingdom.
  • the Antennapedia and Ultrabithorax proteins of Drosophila melanogaster are prototypical homeodomain proteins (Pabo, CO. and R.T. Sauer (1992) Annu. Rev. Biochem. 61:1053-1095).
  • the zinc finger motif which binds zinc ions, generally contains tandem repeats of about 30 amino acids consisting of periodically spaced cysteine and histidine residues. Examples of this sequence pattern, designated C2H2 and C3HC4 ("RING" finger), have been described (Lewin, supra). Zinc finger proteins each contain an ⁇ heUx and an antiparallel ⁇ sheet whose proximity and conformation are maintained by the zinc ion. Contact with DNA is made by the arginine prece ding the ⁇ heUx and by the second, third, and sixth residues of the ⁇ heUx. Variants of the zinc finger motif include poorly defined cysteine-rich motifs which bind zinc or other metal ions. These motifs may not contain histidine residues and are generally nonrepetitive.
  • the leucine zipper motif comprises a stretch of amino acids rich in leucine which can form an amphipafhic ⁇ helix. This structure provides the basis for dimerization of two leucine zipper proteins.
  • the region adjacent to the leucine zipper is usually basic, and upon protein dimerization, is optimally positioned for binding to the major groove. Proteins containing such motifs are generally referred to as bZIP transcription factors.
  • the helix-loop-heUx motif (HLH) consists of a short ⁇ heUx connected by a loop to a longer 5 cc heUx.
  • the loop is flexible and allows the two helices to fold back against each other and to bind to DNA.
  • the transcription factor Myc contains a prototypical HLH motif.
  • Malignant cell growth may result from either excessive expression of tumor promoting genes or insufficient expression of tumor suppressor genes (Cleary, M.L. (1992) Cancer Surv. 15:89-104).
  • Chromosomal translocations may also produce chimeric loci which fuse the coding sequence of one gene with the regulatory regions of a second unrelated gene. Such an arrangement Ukely results in 5 inappropriate gene transcription, potentially contributing to malignancy.
  • the immune system responds to infection or trauma by activating a cascade of events that coordinate the progressive selection, amplification, and mobilization of cellular defense mechanisms.
  • a complex and balanced program of gene activation and repression is involved in this process.
  • hyperactivity of the immune system as a result of improper or insufficient o regulation of gene expression may result in considerable tissue or organ damage. This damage is well documented in immunological responses associated with arthritis, allergens, heart attack, stroke, and infections (Isselbacher, KJ. et al. (1996) Harrison's Principles of Internal Medicine, 13/e, McGraw Hill, Inc. and Teton Data Systems Software).
  • SEQ ID NO:46, SEQ ID NO:47, and SEQ ID NO:48 encode, for example, cell membrane molecules.
  • Eukaryotic cells are surrounded by plasma membranes which enclose the cell and maintain an 5 environment inside the cell that is distinct from its surroundings.
  • eukaryotic organisms are distinct from prokaryotes in possessing many intracellular organelle and vesicle structures. Many of the metabolic reactions which distinguish eukaryotic biochemistry from prokaryotic biochemistry take place within these structures.
  • the plasma membrane and the membranes surrounding organeUes and vesicles are composed of phosphoglycerides, fatty acids, cholesterol, phosphoUpids, glycolipids, 0 proteoglycans, and proteins. These components confer identity and functionaUty to the membranes with which they associate. Integral Membrane Proteins
  • TM proteins transmembrane proteins
  • TM domains are 5 typically comprised of 15 to 25 hydrophobic amino acids which are predicted to adopt an ⁇ -helical conformation.
  • TM proteins are classified as bitopic (Types I and II) and polytopic (Types III and IV) (Singer, S.J. (1990) Annu. Rev. Cell Biol. 6:247-296). Bitopic proteins span the membrane once while polytopic proteins contain multiple membrane-spanning segments.
  • TM proteins function as cell-surface receptors, receptor-interacting proteins, transporters of ions or metabolites, ion channels, o cell anchoring proteins, and cell type-specific surface antigens.
  • MPs membrane proteins
  • PDZ domains KDEL, RGD, NGR, and GSL sequence motifs
  • vWFA von Willebrand factor A
  • EGF-like domains EGF-like domains.
  • RGD, NGR, and GSL motif-containing peptides have been used as drug delivery agents in targeted cancer 5 treatment of tumor vasculature (Aap, W. et al. (1998) Science 279:377-380).
  • MPs may also contain amino acid sequence motifs, such as the carbohydrate recognition domain (CRD), that mediate interactions with extracellular or intracellular molecules.
  • CCD carbohydrate recognition domain
  • GPCR G-protein coupled receptors
  • GPCRs include receptors for biogenic amines, lipid mediators of inflammation, peptide hormones, and sensory signal mediators.
  • the structure of these highly-conserved receptors consists of seven hydrophobic transmembrane regions, an extracellular N-terminus, and a cytoplasmic C-terminus. Three extracellular loops alternate with three intracellular loops to link the seven transmembrane regions. Cysteine disulfide bridges connect the second and 5 third extracellular loops.
  • the most conserved regions of GPCRs are the transmembrane regions and the first two cytoplasmic loops.
  • GPCR-encoding genes A conserved, acidic- g-aromatic residue triplet present in the second cytoplasmic loop may interact with G proteins.
  • a GPCR consensus pattern is characteristic of most proteins belonging to this superfamily (ExPASy PROSITE document PS00237; and Watson, S. and S. Akinstall (1994) The G-protein Linked Receptor Facts Book. Academic Press, San Diego CA, pp. 2-6). Mutations and changes in transcriptional activation of GPCR-encoding genes have been associated with neurological disorders such as schizophrenia, Parkinson's disease, Azheimer' s disease, drug addiction, and feeding disorders. Scavenger Receptors
  • Macrophage scavenger receptors with broad ligand specificity may participate in the binding of low density Upoproteins (LDL) and foreign antigens.
  • Scavenger receptors types I and II are trimeric membrane proteins with each subunit containing a small N-terminal intracellular domain, a transmembrane domain, a large extracellular domain, and a C-terminal cysteine-rich domain.
  • the extracellular domain contains a short spacer region, an ⁇ -helical coiled-coil region, and a triple helical collagen-like region.
  • the transmembrane 4 superfamily (TM4SF) or tetraspan family is a multigene family encoding type III integral membrane proteins (Wright, M.D. and M.G. Tomlinson (1994) Immunol. Today 15 :588-594).
  • the TM4SF is comprised of membrane proteins which traverse the cell membrane four times.
  • Members of the TM4SF include platelet and endothelial cell membrane proteins, melanoma-associated antigens, leukocyte surface glycoproteins, colonal carcinoma antigens, tumor-associated antigens, and surface proteins of the schistosome parasites (Jankowski, S.A. (1994) Oncogene 9:1205-1211).
  • Members of the TM4SF share about 25-30% amino acid sequence identity with one another.
  • TM4SF members have been implicated in signal transduction, control of ceU adhesion, regulation of cell growth and proliferation, including development and oncogenesis, and cell motility, including tumor cell metastasis.
  • Expression of TM4SF proteins is associated with a variety of tumors and the level of expression may be altered when cells are growing or activated.
  • Tumor antigens are cell surface molecules that are differentially expressed in tumor cells relative to normal cells. Tumor antigens distinguish tumor cells immunologically from normal cells and provide diagnostic and therapeutic targets for human cancers (Takagi, S. et al. (1995) Int. J. Cancer 61:706-715; Liu, E. et al. (1992) Oncogene 7:1027-1032).
  • Leukocyte Antigens are cell surface molecules that are differentially expressed in tumor cells relative to normal cells. Tumor antigens.
  • cell surface antigens include those identified on leukocytic cells of the immune system. These antigens have been identified using systematic, monoclonal antibody (mAb)-based "shot gun” techniques. These techniques have resulted in the production of hundreds of mAbs directed against unknown cell surface leukocytic antigens. These antigens have been grouped into “clusters of differentiation” based on common immunocytochemical localization patterns in various differentiated and undifferentiated leukocytic cell types. Antigens in a given cluster are presumed to identify a single cell surface protein and are assigned a "cluster of differentiation" or "CD” designation.
  • mAb monoclonal antibody
  • CD antigens Some of the genes encoding proteins identified by CD antigens have been cloned and verified by standard molecular biology techniques. CD antigens have been characterized as both transmembrane proteins and cell surface proteins anchored to the plasma membrane via covalent attachment to fatty acid-containing glycolipids such as glycosylphosphatidylinositol (GPI). (Reviewed in Barclay, A.N. et al. (1995) The Leucocyte Antigen Facts Book. Academic Press, San Diego CA, pp. 17-20.) Ion Channels
  • Ion channels are found in the plasma membranes of virtually every cell in the body.
  • chloride channels mediate a variety of cellular functions including regulation of membrane potentials and abso ⁇ tion and secretion of ions across epithelial membranes.
  • Chloride channels also regulate the pH of organelles such as the Golgi apparatus and endosomes (see, e.g., Greger, R. (1988) Annu. Rev. Physiol. 50:111-122).
  • Electrophysiological and pharmacological properties of chloride channels including ion conductance, current-voltage relationships, and sensitivity to modulators, suggest that different chloride channels exist in muscles, neurons, fibroblasts, epithelial cells, and lymphocytes.
  • ion channels have sites for phosphorylation by one or more protein kinases including protein kinase A, protein kinase C, tyrosine kinase, and casein kinase II, all of which regulate ion channel activity in cells. Inappropriate phosphorylation of proteins in cells has been Unked to changes in cell cycle progression and cell differentiation. Changes in the cell cycle have been linked to induction of apoptosis or cancer. Changes in cell differentiation have been linked to diseases and disorders of the reproductive system, immune system, skeletal muscle, and other organ systems. Proton Pumps
  • Proton ATPases comprise a large class of membrane proteins that use the energy of ATP hydrolysis to generate an electrochemical proton gradient across a membrane. The resultant gradient may be used to transport other ions across the membrane (Na + , K + , or Cl ' ) or to maintain organelle pH.
  • Proton ATPases are further subdivided into the mitochondrial F- ATPases, the plasma membrane ATPases, and the vacuolar ATPases. The vacuolar ATPases establish and maintain an acidic pH within various organelles involved in the processes of endocytosis and exocytosis (Mellman, I. et al. (1986) Annu. Rev. Biochem. 55:663-700).
  • TAP transporter Another type of peptide transporter, the TAP transporter, is a heterodimer consisting of TAP 1 and TAP 2 and is associated with antigen processing. Peptide antigens are transported across the membrane of the endoplasmic reticulum by l o TAP so they can be expressed on the cell surface in association with MHC molecules.
  • TAP protein consists of multiple hydrophobic membrane spanning segments and a highly conserved ATP-binding cassette (Boll, M. et al. (1996) Proc. Natl.
  • Pathogenic microorganisms such as he ⁇ es simplex virus, may encode inhibitors of TAP-mediated peptide transport in order to evade immune surveillance (Marusina, K. and J.J Manaco (1996) Curr. Opin.
  • ABSC ATP-binding cassette
  • the ATP-binding cassette (ABC) transporters also called the "traffic ATPases”, comprise a superfamily of membrane proteins that mediate transport and channel functions in prokaryotes and eukaryotes (Higgins, CF. (1992) Annu. Rev. Cell Biol. 8:67-113). ABC proteins share a similar
  • ABC transporter genes are associated with various disorders, such as hyperbilirubinemia II/Dubin- Johnson syndrome, recessive Stargardt's disease, X-linked adrenoleukodystrophy, multidrug resistance, celiac disease, and cystic fibrosis.
  • Membrane anchors are covalently joined to a protein post-ttanslationally and include such moieties as prenyl, myristyl, and glycosylphosphatidyl inositol groups.
  • 3 o anchored proteins is important for their function in processes such as receptor-mediated signal transduction. For example, prenylation of Ras is required for its localization to the plasma membrane and for its normal and oncogenic functions in signal transduction.
  • Cells communicate with one another through the secretion and uptake of protein signaling molecules.
  • the uptake of proteins into the cell is achieved by the endocytic pathway, in which the interaction of extracellular signaling molecules with plasma membrane receptors results in the formation of plasma membrane-derived vesicles that enclose and ttansport the molecules into the cytosol. These transport vesicles fuse with and mature into endosomal and lysosomal (digestive) 5 compartments.
  • the secretion of proteins from the cell is achieved by exocytosis, in which molecules inside of the cell proceed through the secretory pathway. In this pathway, molecules transit from the ER to the Golgi apparatus and finally to the plasma membrane, where they are secreted from the cell.
  • vesicles form at the transitional endoplasmic reticulum o (tER), the rim of Golgi cisternae, the face of the Trans-Golgi Network (TGN), the plasma membrane
  • vesicle formation occurs when a region of membrane buds off from the donor organelle.
  • the membrane-bound vesicle contains proteins to be transported and is surrounded by a proteinaceous coat, the components of which are recruited from the cytosol.
  • Two different classes of coat protein have been identified. Clathrin coats form on 5 vesicles derived from the TGN and PM, whereas coatomer (COP) coats form on vesicles derived from the ER and Golgi.
  • COPI involved in retrograde traffic through the Golgi and from the Golgi to the ER
  • COPII involved in anterograde traffic from the ER to the Golgi
  • adapter proteins bring vesicle cargo and coat proteins o together at the surface of the budding membrane.
  • Adapter protein- 1 and -2 select cargo from the
  • TGN and plasma membrane are based on molecular information encoded on the cytoplasmic tail of integral membrane cargo proteins.
  • Adapter proteins also recruit clathrin to the bud site.
  • Clathrin is a protein complex consisting of three large and three small polypeptide chains arranged in a three-legged structure called a triskelion. Multiple triskelions and other coat proteins 5 appear to self -assemble on the membrane to form a coated pit. This assembly process may serve to deform the membrane into a budding vesicle.
  • GTP-bound ADP-ribosylation factor (Arf) is also inco ⁇ orated into the coated assembly.
  • Another small G-protein, dynamin forms a ring complex around the neck of the forming vesicle and may provide the mechanochemical force to seal the bud, thereby releasing the vesicle.
  • the coated vesicle complex is then transported through the cytosol. o During the transport process, Arf-bound GTP is hydrolyzed to GDP, and the coat dissociates from the transport vesicle (West, M.A. et al. (1997) J. Cell Biol. 138:1239-1254).
  • the coat protein is assembled from cytosoUc precursor molecules at specific budding regions on the organelle.
  • the COP coat consists of two 5 major components, a G-protein (Arf or Sar) and coat protomer (coatomer).
  • Coatomer is an equimolar complex of seven proteins, termed alpha-, beta-, beta'-, gamma-, delta-, epsilon- and zeta-COP.
  • the coatomer complex binds to dilysine motifs contained on the cytoplasmic tails of integral membrane proteins.
  • the p24 family of type I membrane proteins represent the major membrane proteins of COPI vesicles (Harter, C and F.T. Wieland (1998) Proc. Natl. Acad. Sci. USA 95:11649-11654).
  • SEQ ID NO:54, SEQ ID NO:55, SEQ ID NO:56, SEQ ID NO:57, SEQ ID NO:58, SEQ ID NO:59, SEQ ID NO:60, SEQ ID NO:61, SEQ ID NO:62, and SEQ ID NO:63 encode, for example, organelle associated molecules.
  • Eukaryotic cells are organized into various cellular organelles which has the effect of separating specific molecules and their functions from one another and from the cytosol. Within the cell, various membrane structures surround and define these organelles while allowing them to interact with one another and the cell environment through both active and passive transport processes. Important cell organelles include the nucleus, the Golgi apparatus, the endoplasmic reticulum, mitochondria, peroxisomes, lysosomes, endosomes, and secretory vesicles. Nucleus
  • the cell nucleus contains all of the genetic information of the cell in the form of DNA, and the components and machinery necessary for replication of DNA and for transcription of DNA into RNA.
  • DNA is organized into compact structures in the nucleus by interactions with various DNA-binding proteins such as histones and non-histone chromosomal proteins.
  • DNA-specific nucleases, DNAses partially degrade these compacted structures prior to DNA replication or transcription.
  • DNA replication takes place with the aid of DNA helicases which unwind the double-stranded DNA helix, and DNA polymerases that duplicate the separated DNA strands.
  • Transcriptional regulatory proteins are essential for the control of gene expression. Some of these proteins function as transcription factors that initiate, activate, repress, or terminate gene transcription. Transcription factors generally bind to the promoter, enhancer, and upstream regulatory regions of a gene in a sequence-specific manner, although some factors bind regulatory elements within or downstream of a gene's coding region. Transcription factors may bind to a specific region of DNA singly or as a complex with other accessory factors. (Reviewed in Lewin, B. (1990) Genes IV, Oxford University Press, New York NY, and Cell Press, Cambridge MA, pp. 554-570.) Many transcription factors inco ⁇ orate DNA-binding structural motifs which comprise either ⁇ helices or ⁇ sheets that bind to the major groove of DNA.
  • helix-turn-heUx helix-turn-heUx
  • zinc finger helix-turn-heUx
  • leucine zipper helix-loop-helix. Proteins containing these motifs may act alone as monomers, or they may form homo- or heterodimers that interact with DNA.
  • Many neoplastic disorders in humans can be attributed to inappropriate gene expression. Malignant cell growth may result from either excessive expression of tumor promoting genes or insufficient expression of tumor suppressor genes (Cleary, MX. (1992) Cancer Surv. 15:89-104).
  • Chromosomal translocations may also produce chimeric loci which fuse the coding sequence of one gene with the regulatory regions of a second unrelated gene. Such an arrangement Ukely results in inappropriate gene transcription, potentially contributing to malignancy.
  • the immune system responds to infection or trauma by activating a cascade of events that coordinate the progressive selection, amplification, and mobilization of cellular defense mechanisms.
  • a complex and balanced program of gene activation and repression is involved in this process.
  • hyperactivity of the immune system as a result of improper or insufficient regulation of gene expression may result in considerable tissue or organ damage. This damage is well documented in immunological responses associated with arthritis, allergens, heart attack, stroke, and infections (Isselbacher, KJ. et al. (1996) Harrison's Principles of Internal Medicine. 13/e, McGraw Hill, Inc. and Teton Data Systems Software).
  • RNA polymerase II transcribes genes that will be translated into proteins.
  • the primary transcript of RNA polymerase II is called heterogenous nuclear RNA (nnRNA), and must be further processed by splicing to remove non-coding sequences called introns.
  • nnRNA heterogenous nuclear RNA
  • RNA splicing is mediated by small nuclear ribonucleoprotein complexes, or snRNPs, producing mature messenger RNA (mRNA) which is then transported out of the nucleus for translation into proteins.
  • the nucleolus is a highly organized subcompartment in the nucleus that contains high concentrations of RNA and proteins and functions mainly in ribosomal RNA synthesis and assembly (Alberts, et al. supra, pp. 379-382).
  • Ribosomal RNA is a structural RNA that is complexed with proteins to form ribonucleoprotein structures called ribosomes. Ribosomes provide the platform on which protein synthesis takes place.
  • Ribosomes are assembled in the nucleolus initially from a large, 45S rRNA combined with a variety of proteins imported from the cytoplasm, as well as smaller, 5S rRNAs. Later processing of the immature ribosome results in formation of smaller ribosomal subunits which are transported from the nucleolus to the cytoplasm where they are assembled into functional ribosomes. Endoplasmic Reticulum
  • proteins are synthesized within the endoplasmic reticulum (ER), deUvered from the ER to the Golgi apparatus for post-translational processing and sorting, and transported from the 5 Golgi to specific intracellular and extracellular destinations. Synthesis of integral membrane proteins, secreted proteins, and proteins destined for the lumen of a particular organelle occurs on the rough endoplasmic reticulum (ER).
  • the rough ER is so named because of the rough appearance in electron micrographs imparted by the attached ribosomes on which protein synthesis proceeds.
  • Protein destined for the ER actually begins in the cytosol with the synthesis of a specific signal 0 peptide which directs the growing polypeptide and its attached ribosome to the ER membrane where the signal peptide is removed and protein synthesis is completed.
  • Soluble proteins destined for the ER lumen, for secretion, or for transport to the lumen of other organelles pass completely into the ER lumen.
  • Transmembrane proteins destined for the ER or for other cell membranes are translocated across the ER membrane but remain anchored in the lipid bilayer of the membrane by one or more 5 membrane-spanning ⁇ -helical regions.
  • Translocated polypeptide chains destined for other organelles or for secretion also fold and assemble in the ER lumen with the aid of certain "resident" ER proteins.
  • Protein folding in the ER is aided by two principal types of protein isomerases, protein disulfide isomerase (PDI), and peptidyl- prolyl isomerase (PPI).
  • PDI protein disulfide isomerase
  • PPI peptidyl- prolyl isomerase
  • PPI peptidyl- prolyl isomerase
  • PPI an enzyme that catalyzes the isomerization of certain proline imide bonds in oligopeptides and proteins, is considered to govern one of the rate limiting steps in the folding of many proteins to their final functional conformation.
  • the cyclophilins represent a major class of PPI that was originally identified as the major receptor for the immunosuppressive drug cyclosporin A (Handschumacher, R.E. et al. (1984) Science 226:544-547).
  • 5 Molecular "chaperones” such as BiP (binding protein) in the ER recognize incorrectly folded proteins as well as proteins not yet folded into their final form and bind to them, both to prevent improper aggregation between them, and to promote proper folding.
  • the Golgi apparatus is a complex structure that lies adjacent to the ER in eukaryotic cells and serves primarily as a sorting and dispatching station for products of the ER (Aberts, et al. supra, pp. 600-610). Additional posttranslational processing, principally additional glycosylation, also occurs in the Golgi. Indeed, the Golgi is a major site of carbohydrate synthesis, including most of the glycosaminoglycans of the extracellular matrix. N-linked oUgosaccharides, added to proteins in the ER, are also further modified in the Golgi by the addition of more sugar residues to form complex N- linked oligosaccharides.
  • O-linked glycosylation of proteins also occurs in the Golgi by the 5 addition of N-acetylgalactosamine to the hydroxyl group of a serine or threonine residue followed by the sequential addition of other sugar residues to the first. This process is catalyzed by a series of glycosyltransferases each specific for a particular donor sugar nucleotide and acceptor molecule (Lodish, H. et al. (1995) Molecular Cell Biology, W.H. Freeman and Co., New York NY, pp.700- 708). In many cases, both N- and O-linked oUgosaccharides appear to be required for the secretion of o proteins or the movement of plasma membrane glycoproteins to the cell surface.
  • the terminal compartment of the Golgi is the Trans-Golgi Network (TGN), where both membrane and lumenal proteins are sorted for their final destination.
  • TGN Trans-Golgi Network
  • Other transport vesicles bud off containing proteins destined for the plasma membrane, such as receptors, adhesion 5 molecules, and ion channels, and secretory proteins, such as hormones, neurotransmitters, and digestive enzymes.
  • the vacuole system is a collection of membrane bound compartments in eukaryotic cells that functions in the processes of endocytosis and exocytosis. They include phagosomes, lysosomes, o endosomes, and secretory vesicles. Endocytosis is the process in cells of internaUzing nutrients, solutes or small particles (pinocytosis) or large particles such as internaUzed receptors, viruses, bacteria, or bacterial toxins (phagocytosis). Exocytosis is the process of transporting molecules to the cell surface. It faciUtates placement or localization of membrane-bound receptors or other membrane proteins and secretion of hormones, neurotransmitters, digestive enzymes, wastes, etc.
  • vacuoles A common property of all of these vacuoles is an acidic pH environment ranging from approximately pH 4.5-5.0. This acidity is maintained by the presence of a proton ATPase that uses the energy of ATP hydrolysis to generate an electrochemical proton gradient across a membrane (Mellman, I. et al. (1986) Annu. Rev. Biochem. 55:663-700).
  • Eukaryotic vacuolar proton ATPase (vp-ATPase) is a multimeric enzyme composed of 3-10 different subunits. One of these subunits is a highly o hydrophobic polypeptide of approximately 16 kDa that is similar to the proteoUpid component of vp-
  • Lysosomes Lysosomes are membranous vesicles containing various hydrolytic enzymes used for the controlled intracellular digestion of macromolecules.
  • Lysosomes contain some 40 types of enzymes including proteases, nucleases, glycosidases, Upases, phosphoUpases, phosphatases, and sulfatases, all of which are acid hydrolases that function at a pH of about 5. Lysosomes are surrounded by a unique 5 membrane containing ttansport proteins that allow the final products of macromolecule degradation, such as sugars, amino acids, and nucleotides, to be transported to the cytosol where they may be either excreted or reutilized by the cell. A vp-ATPase, such as that described above, maintains the acidic environment necessary for hydrolytic activity (Aberts, supra, pp. 610-611).
  • Endosomes o Endosomes are another type of acidic vacuole that is used to transport substances from the cell surface to the interior of the cell in the process of endocytosis. Like lysosomes, endosomes have an acidic environment provided by a vp-ATPase (Alberts et al. supra, pp. 61 -618). Two types of endosomes are apparent based on tracer uptake studies that distinguish their time of formation in the cell and their cellular location. Early endosomes are found near the plasma membrane and appear to 5 function primarily in the recycling of intemaUzed receptors back to the cell surface.
  • Late endosomes appear later in the endocytic process close to the Golgi apparatus and the nucleus, and appear to be associated with delivery of endocytosed material to lysosomes or to the TGN where they may be recycled.
  • Specific proteins are associated with particular transport vesicles and their target compartments that may provide selectivity in targeting vesicles to their proper compartments.
  • a o cytosoUc prenylated GTP-binding protein, Rab is one such protein.
  • Rabs 4, 5, and 11 are associated with the early endosome, whereas Rabs 7 and 9 associate with the late endosome.
  • Mitochondria are oval-shaped organelles comprising an outer membrane, a tightly folded inner membrane, an intermembrane space between the outer and inner membranes, and a matrix 5 inside the inner membrane.
  • the outer membrane contains many porin molecules that allow ions and charged molecules to enter the intermembrane space, while the inner membrane contains a variety of transport proteins that transfer only selected molecules. Mitochondria are the primary sites of energy production in cells.
  • Glucose is initially converted 0 to pyruvate in the cytoplasm.
  • Fatty acids and pyruvate are transported to the mitochondria for complete oxidation to C0 2 coupled by enzymes to the transport of electrons from NADH and FADH 2 to oxygen and to the synthesis of ATP (oxidative phosphorylation) from ADP and P ; .
  • Pyruvate is transported into the mitochondria and converted to acetyl-CoA for oxidation via the citric acid cycle, involving pyruvate dehydrogenase components, dihydrolipoyl transacetylase, and dihydrolipoyl dehydrogenase.
  • Enzymes involved in the citric acid cycle include: citrate synthetase, aconitases, isocitrate dehydrogenase, alpha-ketoglutarate dehydrogenase complex including transsuccinylases, succinyl CoA synthetase, succinate dehydrogenase, fumarases, and malate dehydrogenase.
  • Acetyl CoA is oxidized to C0 2 with concomitant formation of NADH, FADH 2 , and GTP.
  • oxidative phosphorylation the transfer of electrons from NADH and FADH 2 to oxygen by dehydrogenases is coupled to the synthesis of ATP from ADP and P ; by the F ⁇ F ⁇ ATPase complex in the mitochondrial inner membrane.
  • Enzyme complexes responsible for electron transport and ATP synthesis include the FJ ⁇ ATPase complex, ubiquinone(CoQ)-cytochrome c reductase, ubiquinone reductase, cytochrome b, cytochrome c 1( FeS protein, and cytochrome c oxidase.
  • Peroxisomes include the FJ ⁇ ATPase complex, ubiquinone(CoQ)-cytochrome c reductase, ubiquinone reductase, cytochrome b, cytochrome c
  • Peroxisomes like mitochondria, are a major site of oxygen utilization. They contain one or more enzymes, such as catalase and urate oxidase, that use molecular oxygen to remove hydrogen atoms from specific organic substrates in an oxidative reaction that produces hydrogen peroxide (Aberts, supra, pp. 574-577). Catalase oxidizes a variety of substrates including phenols, formic acid, formaldehyde, and alcohol and is important in peroxisomes of liver and kidney cells for detoxifying various toxic molecules that enter the bloodstream.
  • enzymes such as catalase and urate oxidase, that use molecular oxygen to remove hydrogen atoms from specific organic substrates in an oxidative reaction that produces hydrogen peroxide (Aberts, supra, pp. 574-577).
  • Catalase oxidizes a variety of substrates including phenols, formic acid, formaldehyde, and alcohol and is important in peroxisomes of liver and kidney cells for detoxifying various toxic molecules that enter
  • peroxisome assembly factor- 1 Another major function of oxidative reactions in peroxisomes is the breakdown of fatty acids in a process called ⁇ oxidation, ⁇ oxidation results in shortening of the alkyl chain of fatty acids by blocks of two carbon atoms that are converted to acetyl CoA and exported to the cytosol for reuse in biosynthetic reactions.
  • peroxisomes import their proteins from the cytosol using a specific signal sequence located near the C-terminus of the protein. The importance of this import process is evident in the inherited human disease Zellweger syndrome, in which a defect in importing proteins into perixosomes leads to a perixosomal deficiency resulting in severe abnormalities in the brain, liver, and kidneys, and death soon after birth.
  • peroxisome assembly factor- 1 Another major function of oxidative reactions in peroxisomes is the breakdown of fatty acids in a process called ⁇ oxidation, ⁇ oxidation results in shortening of the alkyl chain of fatty acids
  • the present invention relates to nucleic acid sequences comprising human diagnostic and therapeutic polynucleotides (dithp) as presented in the Sequence Listing. Some of the dithp uniquely identify genes encoding human structural, functional, and regulatory molecules.
  • the invention provides an isolated polynucleotide comprising a polynucleotide sequence selected from the group consisting of a) a polynucleotide sequence selected from the group consisting of SEQ ID NO: 1-71 ; b) a naturally occurring polynucleotide sequence having at least 90% sequence identity to a polynucleotide sequence selected from the group consisting of SEQ ID NO: 1-71 ; c) a polynucleotide sequence complementary to a); d) a polynucleotide sequence complementary to b); and e) an RNA equivalent of a) through d).
  • the polynucleotide comprises a polynucleotide sequence selected from the group consisting of SEQ ID NO: 1-71.
  • the polynucleotide comprises at least 60 contiguous nucleotides of a polynucleotide sequence selected from the group consisting of a) a polynucleotide sequence selected from the group consisting of SEQ ID NO: 1 -71 ; b) a naturally occurring polynucleotide sequence having at least 90% sequence identity to a polynucleotide sequence selected from the group consisting of SEQ ID NO: 1-71 ; c) a polynucleotide sequence complementary to a); d) a polynucleotide sequence complementary to b); and e) an RNA equivalent of a) through d).
  • the invention further provides a composition for the detection of expression of human diagnostic and therapeutic polynucleotides, comprising at least one isolated polynucleotide comprising a polynucleotide sequence selected from the group consisting of a) a polynucleotide sequence selected from the group consisting of SEQ ID NO: 1-71 ; b) a naturally occurring polynucleotide sequence having at least 90% sequence identity to a polynucleotide sequence selected from the group consisting of SEQ ID NO: 1-71 ; c) a polynucleotide sequence complementary to a); d) a polynucleotide sequence complementary to b); and e) an RNA equivalent of a) through d); and a detectable label.
  • a composition for the detection of expression of human diagnostic and therapeutic polynucleotides comprising at least one isolated polynucleotide comprising a polynucleotide sequence selected from the group consisting of a) a polyn
  • the invention also provides a method for detecting a target polynucleotide in a sample, said target polynucleotide comprising a polynucleotide sequence selected from the group consisting of a) a polynucleotide sequence selected from the group consisting of SEQ ID NO: 1-71; b) a naturally occurring polynucleotide sequence having at least 90% sequence identity to a polynucleotide sequence selected from the group consisting of SEQ ID NO: 1-71 ; c) a polynucleotide sequence complementary to a); d) a polynucleotide sequence complementary to b); and e) an RNA equivalent of a) through d).
  • the method comprises a) hybridizing the sample with a probe comprising at least 20 contiguous nucleotides comprising a sequence complementary to said target polynucleotide in the sample, and which probe specifically hybridizes to said target polynucleotide, under conditions whereby a hybridization complex is formed between said probe and said target polynucleotide, and b) detecting the presence or absence of said hybridization complex, and, optionally, if present, the amount thereof.
  • the probe comprises at least 30 contiguous nucleotides.
  • the probe comprises at least 60 contiguous nucleotides.
  • the invention further provides a recombinant polynucleotide comprising a promoter sequence operably Unked to an isolated polynucleotide comprising a polynucleotide sequence selected from the group consisting of a) a polynucleotide sequence selected from the group consisting of SEQ ID NO:l- 71 ; b) a naturally occurring polynucleotide sequence having at least 90% sequence identity to a polynucleotide sequence selected from the group consisting of SEQ ID NO: 1-71 ; c) a polynucleotide 5 sequence complementary to a); d) a polynucleotide sequence complementary to b); and e) an RNA equivalent of a) through d).
  • the invention provides a cell transformed with the recombinant polynucleotide.
  • the invention provides a transgenic organism comprising the recombinant polynucleotide.
  • the invention provides a method for producing a human diagnostic and therapeutic polypeptide, the method comprising a) culturing a o cell under conditions suitable for expression of the human diagnostic and therapeutic polypeptide, wherein said cell is transformed with the recombinant polynucleotide, and b) recovering the human diagnostic and therapeutic polypeptide so expressed.
  • the invention also provides a purified human diagnostic and therapeutic polypeptide (DITHP) encoded by at least one polynucleotide comprising a polynucleotide sequence selected from the group 5 consisting of SEQ ID NO:l-71. Additionally, the invention provides an isolated antibody which specifically binds to the human diagnostic and therapeutic polypeptide.
  • DITHP human diagnostic and therapeutic polypeptide
  • the invention further provides a method of identifying a test compound which specifically binds to the human diagnostic and therapeutic polypeptide, the method comprising the steps of a) providing a test compound; b) combining the human diagnostic and therapeutic polypeptide with the test compound for a sufficient time and o under suitable conditions for binding; and c) detecting binding of the human diagnostic and therapeutic polypeptide to the test compound, thereby identifying the test compound which specifically binds the human diagnostic and therapeutic polypeptide.
  • the invention further provides a microarray wherein at least one element of the microarray is an isolated polynucleotide comprising at least 60 contiguous nucleotides of a polynucleotide comprising 5 a polynucleotide sequence selected from the group consisting of a) a polynucleotide sequence selected from the group consisting of SEQ ID NO: 1 -71 ; b) a naturally occurring polynucleotide sequence having at least 90% sequence identity to a polynucleotide sequence selected from the group consisting of SEQ ID NO:l-71 ; c) a polynucleotide sequence complementary to a); d) a polynucleotide sequence complementary to b); and e) an RNA equivalent of a) through d).
  • the invention also provides a method o for generating a transcript image of a sample which contains polynucleotides.
  • the method comprises a) labeling the polynucleotides of the sample, b) contacting the elements of the microarray with the labeled polynucleotides of the sample under conditions suitable for the formation of a hybridization complex, and c) quantifying the expression of the polynucleotides in the sample.
  • the invention provides a method for screening a compound for effectiveness in altering expression of a target polynucleotide, wherein said target polynucleotide comprises a polynucleotide sequence selected from the group consisting of a) a polynucleotide sequence selected from the group consisting of SEQ ID NO: 1-71 ; b) a naturally occurring polynucleotide sequence having at least 90% sequence identity to a polynucleotide sequence selected from the group consisting of SEQ 5 ID NO:l-71 ; c) a polynucleotide sequence complementary to a); d) a polynucleotide sequence complementary to b); and e) an RNA equivalent of a) through d).
  • the method comprises a) exposing a sample comprising the target polynucleotide to a compound, and b) detecting altered expression of the target polynucleotide.
  • the invention further provides a method for assessing toxicity of a test compound, said method 0 comprising a) treating a biological sample containing nucleic acids with the test compound; b) hybridizing the nucleic acids of the treated biological sample with a probe comprising at least 20 contiguous nucleotides of a polynucleotide comprising a polynucleotide sequence selected from the group consisting of i) a polynucleotide sequence selected from the group consisting of SEQ ID NO:l- 71 ; U) a naturally occurring polynucleotide sequence having at least 90% sequence identity to a 5 polynucleotide sequence selected from the group consisting of SEQ ID NO : 1 -71 ; Ui) a polynucleotide sequence complementary to i), iv) a polynucleotide sequence complementary to u), and v) an RNA equivalent of i)-iv).
  • Hybridization occurs under conditions whereby a specific hybridization complex is formed between said probe and a target polynucleotide in the biological sample, said target polynucleotide comprising a polynucleotide sequence selected from the group consisting of i) a o polynucleotide sequence selected from the group consisting of SEQ ID NO: 1 -71 ; ii) a naturally occurring polynucleotide sequence having at least 90% sequence identity to a polynucleotide sequence selected from the group consisting of SEQ ID NO: 1-71; iii) a polynucleotide sequence complementary to i), iv) a polynucleotide sequence complementary to ii), and v) an RNA equivalent of i)-iv), and alternatively, the target polynucleotide comprises a fragment of a polynucleotide sequence 5 selected from the group consisting of i-v above; c) quantifying the amount of hybridization complex
  • Table 1 shows the sequence identification numbers (SEQ ID NO:s) and template identification numbers (template IDs) corresponding to the polynucleotides of the present invention, along with their GenBank hits (GI Numbers), probabiUty scores, and functional annotations corresponding to the GenBank hits.
  • Table 2 shows the sequence identification numbers (SEQ ID NO:s) and template identification numbers (template IDs) corresponding to the polynucleotides of the present invention, along with polynucleotide segments of each template sequence as defined by the indicated "start” and "stop” nucleotide positions. The reading frames of the polynucleotide segments and the Pfam hits, Pfam 5 descriptions, and E- values corresponding to the polypeptide domains encoded by the polynucleotide segments are indicated.
  • Table 3 shows the sequence identification numbers (SEQ ID NO:s) and template identification numbers (template IDs) corresponding to the polynucleotides of the present invention, along with polynucleotide segments of each template sequence as defined by the indicated “start” and “stop” o nucleotide positions.
  • the reading frames of the polynucleotide segments are shown, and the polypeptides encoded by the polynucleotide segments constitute either signal peptide (SP) or transmembrane (TM) domains, as indicated.
  • SP signal peptide
  • TM transmembrane
  • Table 4 shows the sequence identification numbers (SEQ ID NO:s) and template identification numbers (template IDs) corresponding to the polynucleotides of the present invention, along with 5 component sequence identification numbers (component IDs) corresponding to each template.
  • the component sequences, which were used to assemble the template sequences, are defined by the indicated “start” and “stop” nucleotide positions along each template.
  • Table 5 shows the tissue distribution profiles for the templates of the invention.
  • Table 6 summarizes the bioinformatics tools which are useful for analysis of the o polynucleotides of the present invention.
  • the first column of Table 6 Usts analytical tools, programs, and algorithms, the second column provides brief descriptions thereof, the third column presents appropriate references, all of which are inco ⁇ orated by reference herein in their entirety, and the fourth column presents, where appUcable, the scores, probabiUty values, and other parameters used to evaluate the strength of a match between two sequences (the higher the score, the greater the homology between 5 two sequences).
  • dithp refers to a nucleic acid sequence
  • DITHP amino acid sequence encoded by dithp
  • a “full-length” dithp refers to a nucleic o acid sequence containing the entire coding region of a gene endogenously expressed in human tissue.
  • Adjuvants are materials such as Freund's adjuvant, mineral gels (aluminum hydroxide), and surface active substances (lysolecitbin, pluronic polyols, polyanions, peptides, oil emulsions, keyhole limpet hemocyanin, and dinifrophenol) which may be administered to increase a host's immunological response.
  • Alele refers to an alternative form of a nucleic acid sequence. Aleles result from a
  • Mutation a change or an alternative reading of the genetic code.
  • a y given gene may have none, one, or many allelic forms. Mutations which give rise to alleles include deletions, additions, or substitutions of nucleotides. Each of these changes may occur alone, or in combination with the others, one or more times in a given nucleic acid sequence.
  • the present invention encompasses alleUc dithp. 0
  • Amino acid sequence refers to a peptide, a polypeptide, or a protein of either natural or synthetic origin. The amino acid sequence is not Umited to the complete, endogenous amino acid sequence and may be a fragment, epitope, variant, or derivative of a protein expressed by a nucleic acid sequence.
  • PCR polymerase chain reaction
  • Aitibody refers to intact molecules as well as to fragments thereof, such as Fab, F(ab') 2 , and Fv fragments, which are capable of binding the epitopic determinant.
  • Aitibodies that bind DITHP polypeptides can be prepared using intact polypeptides or using fragments containing small peptides of interest as the immunizing antigen.
  • the polypeptide or peptide used to immunize an animal e.g., a o mouse, a rat, or a rabbit
  • an animal e.g., a o mouse, a rat, or a rabbit
  • Commonly used carriers that are chemically coupled to peptides include bovine serum albumin, thyroglobulin, and keyhole Umpet hemocyanin (KLH). The coupled peptide is then used to immunize the animal.
  • Antisense sequence refers to a sequence capable of specifically hybridizing to a target sequence.
  • the antisense sequence may include DNA, RNA, or any nucleic acid mimic or analog such as peptide nucleic acid (PNA); oUgonucleotides having modified backbone Unkages such as phosphorothioates, methylphosphonates, or benzylphosphonates; oUgonucleotides having modified sugar groups such as 2 -methoxyethyl sugars or 2 -methoxyethoxy sugars; or oligonucleotides having modified bases such as 5-methyl cytosine, 2 -deoxyuracil, or 7-deaza-2'-deoxyguanosine.
  • PNA peptide nucleic acid
  • Antisense sequence refers to a sequence capable of specifically hybridizing to a target sequence.
  • the antisense sequence can be DNA, RNA, or any nucleic acid mimic or analog.
  • Antisense technology refers to any technology which reUes on the specific hybridization of an antisense sequence to a target sequence.
  • a “bin” is a portion of computer memory space used by a computer program for storage of data, and bounded in such a manner that data stored in a bin may be retrieved by the program.
  • Bioly active refers to an amino acid sequence having a structural, regulatory, or biochemical function of a naturally occurring amino acid sequence.
  • “Clone joining” is a process for combining gene bins based upon the bins' containing sequence information from the same clone.
  • the sequences may assemble into a primary gene transcript as well as one or more spUce variants.
  • “Complementary” describes the relationship between two single-stranded nucleic acid sequences that anneal by base-pairing (5'-A-G-T-3' pairs with its complement 3'-T-C-A-5').
  • a “component sequence” is a nucleic acid sequence selected by a computer program such as PHRED and used to assemble a consensus or template sequence from one or more component sequences.
  • a "consensus sequence” or “template sequence” is a nucleic acid sequence which has been assembled from overlapping sequences, using a computer program for fragment assembly such as the GEL VIEW fragment assembly system (Genetics Computer Group (GCG), Madison WI) or using a relational database management system (RDMS).
  • GCG Genetics Computer Group
  • RDMS relational database management system
  • Constant amino acid substitutions are those substitutions that, when made, least interfere with the properties of the original protein, i.e., the structure and especially the function of the protein is conserved and not significantly changed by such substitutions.
  • the table below shows amino acids which may be substituted for an original amino acid in a protein and which are regarded as conservative substitutions.
  • Conservative substitutions generally maintain (a) the structure of the polypeptide backbone in the area of the substitution, for example, as a beta sheet or alpha hetical conformation, (b) the charge or hydrophobicity of the molecule at the target site, or (c) the bulk of the side chain.
  • “Deletion” refers to a change in either a nucleic or amino acid sequence in which at least one nucleotide or amino acid residue, respectively, is absent.
  • Derivative refers to the chemical modification of a nucleic acid sequence, such as by replacement of hydrogen by an alkyl, acyl, amino, hydroxyl, or other group.
  • array element refers to a polynucleotide, polypeptide, or other chemical compound having a unique and defined position on a microarray.
  • a "fragment” is a unique portion of dithp or DITHP which is identical in sequence to but shorter in length than the parent sequence.
  • a fragment may comprise up to the entire length of the defined sequence, minus one nucleotide/amino acid residue. For example, a fragment may comprise from 10 to 1000 contiguous amino acid residues or nucleotides.
  • a fragment used as a probe, primer, antigen, therapeutic molecule, or for other pu ⁇ oses may be at least 5, 10, 15, 16, 20, 25, 30, 40, 50, 60, 75, 100, 150, 250 or at least 500 contiguous amino acid residues or nucleotides in length.
  • Fragments may be preferentially selected from certain regions of a molecule.
  • a polypeptide fragment may comprise a certain length of contiguous amino acids selected from the first 250 or 500 amino acids (or first 25% or 50%) of a polypeptide as shown in a certain defined sequence.
  • these lengths are exemplary, and any length that is supported by the specification, including the Sequence Listing and the figures, may be encompassed by the present embodiments.
  • a fragment of dithp comprises a region of unique polynucleotide sequence that specifically identifies dithp, for example, as distinct from any other sequence in the same genome.
  • a fragment of 5 dithp is useful, for example, in hybridization and amplification technologies and in analogous methods that distinguish dithp from related polynucleotide sequences.
  • the precise length of a fragment of dithp and the region of dithp to which the fragment corresponds are routinely determinable by one of ordinary skill in the art based on the intended pu ⁇ ose for the fragment.
  • a fragment of DITHP is encoded by a fragment of dithp.
  • a fragment of DITHP comprises a 0 region of unique amino acid sequence that specifically identifies DITHP.
  • a fragment of DITHP is useful as an immunogenic peptide for the development of antibodies that specifically recognize DITHP.
  • the precise length of a fragment of DITHP and the region of DITHP to which the fragment corresponds are routinely determinable by one of ordinary skill in the art based on the intended pu ⁇ ose for the fragment. 5
  • a "full length" nucleotide sequence is one containing at least a start site for translation to a protein sequence, followed by an open reading frame and a stop site, and encoding a "full length" polypeptide.
  • “Hit” refers to a sequence whose annotation will be used to describe a given template. Criteria for selecting the top hit are as follows: if the template has one or more exact nucleic acid matches, the o top hit is the exact match with highest percent identity. If the template has no exact matches but has significant protein hits, the top hit is the protein hit with the lowest E-value. If the template has no significant protein hits, but does have significant non-exact nucleotide hits, the top hit is the nucleotide hit with the lowest E-value.
  • Homology refers to sequence similarity either between a reference nucleic acid sequence and 5 at least a fragment of a dithp or between a reference amino acid sequence and a fragment of a DITHP.
  • Hybridization refers to the process by which a strand of nucleotides anneals with a complementary strand through base pairing. Specific hybridization is an indication that two nucleic acid sequences share a high degree of identity. Specific hybridization complexes form under defined annealing conditions, and remain hybridized after the "washing" step.
  • the defined hybridization o conditions include the anneahng conditions and the washing step(s), the latter of which is particularly important in determining the stringency of the hybridization process, with more stringent conditions allowing less non-specific binding, i.e., binding between pairs of nucleic acid probes that are not perfectly matched.
  • Permissive conditions for anneahng of nucleic acid sequences are routinely determinable and may be consistent among hybridization experiments, whereas wash conditions may be varied among experiments to achieve the desired stringency.
  • TJ thermal melting point
  • T m is the temperature (under defined ionic strength and pH) at which 50% of the target sequence hybridizes to a perfectly matched probe.
  • High stringency conditions for hybridization between polynucleotides of the present invention include wash conditions of 68°C in the presence of about 0.2 x SSC and about 0.1 % SDS, for 1 hour. Aternatively, temperatures of about 65°C, 60°C, or 55°C may be used. SSC concentration may be varied from about 0.2 to 2 x SSC, with SDS being present at about 0.1 %.
  • blocking reagents 5 are used to block non-specific hybridizatioa Such blocking reagents include, for instance, denatured salmon sperm DNA at about 100-200 ⁇ g/ml. Useful variations on these conditions will be readily apparent to those skilled in the art.
  • Hybridization particularly under high stringency conditions, may be suggestive of evolutionary similarity between the nucleotides. Such similarity is strongly indicative of a similar role for the nucleotides and their resultant proteins. o Other parameters, such as temperature, salt concentration, and detergent concentration may be varied to achieve the desired stringency. Denaturants, such as formamide at a concentration of about 35-50% v/v, may also be used under particular circumstances, such as RNA:DNA hybridizations. Appropriate hybridization conditions are routinely determinable by one of ordinary skill in the art.
  • Immunogenic describes the potential for a natural, recombinant, or synthetic peptide, epitope, 5 polypeptide, or protein to induce antibody production in appropriate animals, cells, or cell lines.
  • “Insertion” or “addition” refers to a change in either a nucleic or amino acid sequence in which at least one nucleotide or residue, respectively, is added to the sequence.
  • LabeleUng refers to the covalent or noncovalent joining of a polynucleotide, polypeptide, or antibody with a reporter molecule capable of producing a detectable or measurable signal.
  • Mcroarray is any arrangement of nucleic acids, amino acids, antibodies, etc., on a substrate.
  • the substrate may be a sohd support such as beads, glass, paper, nitrocellulose, nylon, or an appropriate membrane.
  • Linkers are short stretches of nucleotide sequence which may be added to a vector or a dithp to create restriction endonuclease sites to faciUtate cloning.
  • PolyUnkers are engineered to inco ⁇ orate multiple restriction enzyme sites and to provide for the use of enzymes which leave 5 ' or 3 ' overhangs (e.g., BamHI, EcoRI, and Hindlll) and those which provide blunt ends (e.g., EcoRV, SnaBI, and Stul).
  • Nucleic acid sequence refers to the specific order of nucleotides joined by phosphodiester bonds in a linear, polymeric arrangement. Depending on the number of nucleotides, the nucleic acid sequence can be considered an oUgomer, oligonucleotide, or polynucleotide.
  • the nucleic acid can be DNA, RNA, or any nucleic acid analog, such as PNA, may be of genomic or synthetic origin, may be either double-stranded or single-stranded, and can represent either the sense or antisense 0 (complementary) strand.
  • OUgomer refers to a nucleic acid sequence of at least about 6 nucleotides and as many as about 60 nucleotides, preferably about 15 to 40 nucleotides, and most preferably between about 20 and 30 nucleotides, that may be used in hybridization or ampUfication technologies. OUgomers may be used as, e.g., primers for PCR, and are usually chemically synthesized. 5 "Operably linked" refers to the situation in which a first nucleic acid sequence is placed in a functional relationship with the second nucleic acid sequence. For instance, a promoter is operably Unked to a coding sequence if the promoter affects the transcription or expression of the coding sequence.
  • operably Unked DNA sequences may be in close proximity or contiguous and, where necessary to join two protein coding regions, in the same reading frame.
  • PNA protein nucleic acid
  • PNA refers to a DNA mimic in which nucleotide bases are attached to a pseudopeptide backbone to increase stability.
  • PNA also designated antigene agents, can prevent gene expression by targeting complementary messenger RNA.
  • percent identity and % identity refer to the percentage of residue matches between at least two polynucleotide sequences aUgned using a 5 standardized algorithm. Such an algorithm may insert, in a standardized and reproducible way, gaps in the sequences being compared in order to optimize aUgnment between two sequences, and therefore achieve a more meaningful comparison of the two sequences.
  • Percent identity between polynucleotide sequences may be determined using the default parameters of the CLUSTAL V algorithm as inco ⁇ orated into the MEGALIGN version 3.12e sequence o aUgnment program. This program is part of the LASERGENE software package, a suite of molecular biological analysis programs (DNASTAR, Madison WI). CLUSTAL V is described in Higgins, D.G. and Sha ⁇ , P.M. (1989) CABIOS 5:151-153 and in Higgins, D.G. et al. (1992) CABIOS 8:189-191.
  • BLAST Basic Local AUgnment Search 5 Tool
  • BLAST 2 Sequences can be accessed and used interactively at http://www.ncbi.nlm.nih.gov/gorf/bl2/.
  • the "BLAST 2 Sequences” tool can be used for both blastn and blastp (discussed below).
  • BLAST programs are commonly used with gap and other parameters set to default settings. For example, to compare two nucleotide sequences, one may use blastn with the "BLAST 2 Sequences" tool Version 5 2.0.9 (May-07-1999) set at default parameters.
  • Such default parameters may be, for example:
  • Percent identity may be measured over the length of an entire defined sequence, for example, as 5 defined by a particular SEQ ID number, or may be measured over a shorter length, for example, over the length of a fragment taken from a larger, defined sequence, for instance, a fragment of at least 20, at least 30, at least 40, at least 50, at least 70, at least 100, or at least 200 contiguous nucleotides.
  • Such lengths are exemplary only, and it is understood that any fragment length supported by the sequences shown herein, in figures or Sequence Listings, may be used to describe a length over which percentage o identity may be measured.
  • Nucleic acid sequences that do not show a high degree of identity may nevertheless encode similar amino acid sequences due to the degeneracy of the genetic code. It is understood that changes in nucleic acid sequence can be made using this degeneracy to produce multiple nucleic acid sequences that all encode substantially the same proteia
  • the phrases "percent identity” and "% identity", as appUed to polypeptide sequences refer to the percentage of residue matches between at least two polypeptide sequences aUgned using a standardized algorithm. Methods of polypeptide sequence aUgnment are well-known. Some aUgnment methods take into account conservative amino acid substitutions. Such conservative substitutions, explained in more detail above, generally preserve the hydrophobicity and acidity of the substituted residue, thus preserving the structure (and therefore function) of the folded polypeptide.
  • NCBI BLAST software suite may be used.
  • BLAST 2 Sequences Version 2.0.9 (May-07-1999) with blastp set at default parameters.
  • Such default parameters may be, for example:
  • Percent identity may be measured over the length of an entire defined polypeptide sequence, for example, as defined by a particular SEQ ID number, or may be measured over a shorter length, for example, over the length of a fragment taken from a larger, defined polypeptide sequence, for instance, a fragment of at least 15, at least 20, at least 30, at least 40, at least 50, at least 70 or at least 150 contiguous residues.
  • Such lengths are exemplary only, and it is understood that any fragment length supported by the sequences shown herein, in figures or Sequence Listings, may be used to describe a length over which percentage identity may be measured.
  • Probe refers to dithp or fragments thereof, which are used to detect identical, alletic or related nucleic acid sequences. Probes are isolated oUgonucleotides or polynucleotides attached to a detectable label or reporter molecule. Typical labels include radioactive isotopes, Ugands, chemiluminescent agents, and enzymes.
  • Primer pairs are short nucleic acids, usually DNA oUgonucleotides, which may be 5 annealed to a target polynucleotide by complementary base-pairing. The primer may then be extended along the target DNA strand by a DNA polymerase enzyme. Primer pairs can be used for ampUfication (and identification) of a nucleic acid sequence, e.g., by the polymerase chain reaction (PCR).
  • PCR polymerase chain reaction
  • Probes and primers as used in the present invention typically comprise at least 15 contiguous nucleotides of a known sequence. In order to enhance specificity, longer probes and primers may also 0 be employed, such as probes and primers that comprise at least 20, 30, 40, 50, 60, 70, 80, 90, 100, or at least 150 consecutive nucleotides of the disclosed nucleic acid sequences. Probes and primers may be considerably longer than these examples, and it is understood that any length supported by the specification, including the figures and Sequence Listing, may be used.
  • PCR primer pairs can be derived from a known sequence, for example, by using computer programs intended for that pu ⁇ ose such as Primer o (Version 0.5, 1991 , Whitehead Institute for Biomedical Research, Cambridge MA).
  • OUgonucleotides for use as primers are selected using software known in the art for such pu ⁇ ose. For example, OLIGO 4.06 software is useful for the selection of PCR primer pairs of up to 100 nucleotides each, and for the analysis of oUgonucleotides and larger polynucleotides of up to 5,000 nucleotides from an input polynucleotide sequence of up to 32 kilobases. Similar primer selection 5 programs have inco ⁇ orated additional features for expanded capabiUties.
  • the PrimOU primer selection program (available to the pubUc from the Genome Center at University of Texas South West Medical Center, Dallas TX) is capable of choosing specific primers from megabase sequences and is thus useful for designing primers on a genome-wide scope.
  • the Primer3 primer selection program (available to the pubUc from the Whitehead Institute/MIT Center for Genome Research, o Cambridge MA) allows the user to input a "mispriming Ubrary," in which sequences to avoid as primer binding sites are user-specified. Primer3 is useful, in particular, for the selection of oUgonucleotides for microarrays.
  • the source code for the latter two primer selection programs may also be obtained from their respective sources and modified to meet the user's specific needs.
  • the PrimeGen program (available to the public from the UK Human Genome Mapping Project Resource Centre, Cambridge UK) designs primers based on multiple sequence alignments, thereby allowing selection of primers that hybridize to either the most conserved or least conserved regions of aUgned nucleic acid sequences.
  • this program is useful for identification of both unique and conserved oligonucleotides and polynucleotide fragments.
  • the oUgonucleotides and polynucleotide fragments identified by any of the above selection methods are useful in hybridization technologies, for example, as PCR or sequencing primers, microarray elements, or specific probes to identify fully or partially complementary polynucleotides in a sample of nucleic acids. Methods of oUgonucleotide selection are not Umited to those described above.
  • “Purified” refers to molecules, either polynucleotides or polypeptides that are isolated or separated from their natural environment and are at least 60% free, preferably at least 75% free, and most preferably at least 90% free from other compounds with which they are naturally associated.
  • a "recombinant nucleic acid” is a sequence that is not naturally occurring or has a sequence that is made by an artificial combination of two or more otherwise separated segments of sequence.
  • This artificial combination is often accompUshed by chemical synthesis or, more commonly, by the artificial manipulation of isolated segments of nucleic acids , e. g. , by genetic engineering techniques such as those described in Sambrook, supra.
  • the term recombinant includes nucleic acids that have been altered solely by addition, substitution, or deletion of a portion of the nucleic acid.
  • a recombinant nucleic acid may include a nucleic acid sequence operably Unked to a promoter sequence.
  • Such a recombinant nucleic acid may be part of a vector that is used, for example, to transform a cell.
  • recombinant nucleic acids may be part of a viral vector, e.g., based on a vaccinia virus, that could be use to vaccinate a mammal wherein the recombinant nucleic acid is expressed, inducing a protective immunological response in the mammal.
  • regulatory element refers to a nucleic acid sequence from nontranslated regions of a gene, and includes enhancers, promoters, introns, and 3' untranslated regions, which interact with host proteins to carry out or regulate transcription or translation.
  • RNA equivalent in reference to a DNA sequence, is composed of the same Unear sequence of nucleotides as the reference DNA sequence with the exception that all occurrences of the nitrogenous base thymine are replaced with uracil, and the sugar backbone is composed of ribose instead of deoxyribose.
  • sample is used in its broadest sense.
  • Samples may contain nucleic or amino acids, antibodies, or other materials, and may be derived from any source (e.g., bodily fluids including, but not Umited to, saUva, blood, and urine; chromosome(s), organelles, or membranes isolated from a cell; genomic DNA, RNA, or cDNA in solution or bound to a substrate; and cleared cells or tissues or blots 5 or imprints from such cells or tissues).
  • bodily fluids including, but not Umited to, saUva, blood, and urine
  • chromosome(s), organelles, or membranes isolated from a cell genomic DNA, RNA, or cDNA in solution or bound to a substrate
  • cleared cells or tissues or blots 5 or imprints from such cells or tissues e.g., bodily fluids including, but not Umited to, saUva, blood, and urine; chromosome(s), organelles, or membranes isolated from a cell; genomic DNA, RNA, or
  • Specific binding or “specifically binding” refers to the interaction between a protein or peptide and its agonist, antibody, antagonist, or other binding partner. The interaction is dependent upon the presence of a particular structure of the protein, e.g., the antigenic determinant or epitope, recognized by the binding molecule. For example, if an antibody is specific for epitope "A,” the o presence of a polypeptide containing epitope A, or the presence of free unlabeled A, in a reaction containing free labeled A and the antibody will reduce the amount of labeled A that binds to the antibody.
  • Substitution refers to the replacement of at least one nucleotide or amino acid by a different nucleotide or amino acid.
  • Substrate refers to any suitable rigid or semi-rigid support including, e.g., membranes, filters, chips, sUdes, wafers, fibers, magnetic or nonmagnetic beads, gels, tubing, plates, polymers, microparticles or capillaries.
  • the substrate can have a variety of surface forms, such as wells, trenches, pins, channels and pores, to which polynucleotides or polypeptides are bound.
  • a “transcript image” refers to the collective pattern of gene expression by a particular tissue or o cell type under given conditions at a given time.
  • Transformation refers to a process by which exogenous DNA enters a recipient cell. Transformation may occur under natural or artificial conditions using various methods well known in the art. Transformation may rely on any known method for the insertion of foreign nucleic acid sequences into a prokaryotic or eukaryotic host cell. The method is selected based on the host cell being 5 transformed.
  • Transformants include stably transformed cells in which the inserted DNA is capable of replication either as an autonomously repUcating plasmid or as part of the host chromosome, as well as cells which transiently express inserted DNA or RNA.
  • a "transgenic organism,” as used herein, is any organism, including but not Umited to animals o and plants, in which one or more of the cells of the organism contains heterologous nucleic acid introduced by way of human intervention, such as by transgenic techniques well known in the art.
  • the nucleic acid is introduced into the cell, directly or indirectly by introduction into a precursor of the cell, by way of deUberate genetic manipulation, such as by microinjection or by infection with a recombinant virus.
  • the term genetic manipulation does not include classical cross-breeding, or in vitro fertiUzation, but rather is directed to the introduction of a recombinant DNA molecule.
  • the transgenic organisms contemplated in accordance with the present invention include bacteria, cyanobacteria, fungi, and plants and animals.
  • the isolated DNA of the present invention can be introduced into the host by methods known in the art, for example infection, transfection, transformation or transconjugatioa Techniques for transferring the DNA of the present invention into such organisms are widely known and provided in references such as Sambrook et al. (1989), supra.
  • a “variant" of a particular nucleic acid sequence is defined as a nucleic acid sequence having at least 25% sequence identity to the particular nucleic acid sequence over a certain length of one of the nucleic acid sequences using blastn with the "BLAST 2 Sequences" tool Version 2.0.9 (May-07-1999) set at default parameters.
  • Such a pair of nucleic acids may show, for example, at least 30%, at least 50%, at least 60%, at least 70%, at least 80%, at least 90%, at least 95% or even at least 98% or greater sequence identity over a certain defined length.
  • the variant may result in "conservative" amino acid changes which do not affect structural and/or chemical properties.
  • a variant may be described as, for example, an "allelic” (as defined above), “splice,” “species,” or “polymo ⁇ hic” variant.
  • a spUce variant may have significant identity to a reference molecule, but will generally have a greater or lesser number of polynucleotides due to alternate spUcing of exons during mRNA processing.
  • the corresponding polypeptide may possess additional functional domains or lack domains that are present in the reference molecule.
  • Species variants are polynucleotide sequences that vary from one species to another. The resulting polypeptides generally will have significant amino acid identity relative to each other.
  • a polymo ⁇ hic variant is a variation in the polynucleotide sequence of a particular gene between individuals of a given species.
  • Polymo ⁇ hic variants also may encompass "single nucleotide polymorphisms" (SNPs) in which the polynucleotide sequence varies by one base. The presence of SNPs may be indicative of, for example, a certain population, a disease state, or a propensity for a disease state.
  • variants of the polynucleotides of the present invention may be generated through recombinant methods.
  • One possible method is a DNA shuffling technique such as MOLECULARBREEDING (Maxygen Inc., Santa Clara CA; described in U.S.
  • DNA shuffling is a process by which a Ubrary of gene variants is produced using PCR-mediated recombination of gene fragments.
  • the Ubrary is then subjected to selection or screening procedures that identify those gene variants with the desired properties. These preferred variants may then be pooled and further subjected to recursive rounds of DNA shuffling and selection/screening.
  • genetic diversity is created through "artificial" breeding and rapid molecular evolution. For example, fragments of a single gene containing random point mutations may be recombined, screened, and then reshuffled until the desired properties are optimized. Aternatively, fragments of a given gene may be recombined with fragments of homologous genes in the same gene family, either from the same or different species, thereby maximizing the genetic diversity of multiple naturally occurring genes in a directed and controllable manner.
  • a "variant" of a particular polypeptide sequence is defined as a polypeptide sequence having at least 40% sequence identity to the particular polypeptide sequence over a certain length of one of the polypeptide sequences using blastp with the "BLAST 2 Sequences" tool Version 2.0.9 (May-07- 1999) set at default parameters.
  • Such a pair of polypeptides may show, for example, at least 50%, at least 60%, at least 70%, at least 80%, at least 90%, at least 95%, or at least 98% or greater sequence identity over a certain defined length of one of the polypeptides.
  • cDNA sequences derived from human tissues and cell Unes were aUgned based on nucleotide sequence identity and assembled into "consensus" or "template” sequences which are designated by the template identification numbers (template IDs) in column 2 of Table 1.
  • the sequence identification numbers (SEQ ID NO:s) corresponding to the template IDs are shown in column 1.
  • the template sequences have similarity to GenBank sequences, or "hits," as designated by the GI Numbers in column 3.
  • the statistical probability of each GenBank hit is indicated by a probabiUty score in column 4, and the functional annotation corresponding to each GenBank hit is Usted in column 5.
  • the invention inco ⁇ orates the nucleic acid sequences of these templates as disclosed in the Sequence Listing and the use of these sequences in the diagnosis and treatment of disease states characterized by defects in human molecules.
  • the invention further utiUzes these sequences in hybridization and ampUfication technologies, and in particular, in technologies which assess gene expression patterns correlated with specific cells or tissues and their responses in vivo or in vitro to pharmaceutical agents, toxins, and other treatments. In this manner, the sequences of the present invention are used to develop a transcript image for a particular cell or tissue.
  • cDNA was isolated from Ubraries constructed using RNA derived from normal and diseased human tissues and cell Unes.
  • the human tissues and cell Unes used for cDNA Ubrary construction were selected from a broad range of sources to provide a diverse population of cDNA representative of gene transcription throughout the human body. Descriptions of the human tissues and cell Unes used for cDNA Ubrary construction are provided in the LEFESEQ database (Incyte Genomics, Inc. (Incyte), Palo Ato CA).
  • Human tissues were broadly selected from, for example, cardiovascular, dermatologic, endocrine, gastrointestinal, hematopoietic/immune system, musculoskeletal, neural, reproductive, and urologic sources.
  • Cell Unes used for cDNA library construction were derived from, for example, leukemic cells, teratocarcinomas, neuroepitheUomas, cervical carcinoma, lung fibroblasts, and endotheUal cells.
  • Such cell lines include, for example, THP-1, Jurkat, HUVEC, hNT2, WI38, HeLa, and other cell Unes commonly used and available from pubUc depositories (American Type Culture Collection, Manassas VA).
  • cell Unes Prior to mRNA isolation, cell Unes were untreated, treated with a pharmaceutical agent such as 5 -aza-2'-deoxycytidine, treated with an activating agent such as Upopolysaccharide in the case of leukocytic cell Unes, or, in the case of endotheUal cell lines, subjected to shear stress.
  • a pharmaceutical agent such as 5 -aza-2'-deoxycytidine
  • an activating agent such as Upopolysaccharide in the case of leukocytic cell Unes, or, in the case of endotheUal cell lines, subjected to shear stress.
  • Sequencing of the cDNA Methods for DNA sequencing are well known in the art.
  • Conventional enzymatic methods employ the Klenow fragment of DNA polymerase I, SEQUENASE DNA polymerase (U.S. Biochemical Co ⁇ oration, Cleveland OH), Taq polymerase (PE Biosystems, Foster City CA), thermostable T7 polymerase (Amersham Pharmacia Biotech, Inc. (Amersham Pharmacia Biotech), Piscataway NJ), or combinations of polymerases and proofreading exonucleases such as those found in the ELONGASE ampUfication system (Life Technologies Inc. (Life Technologies), Gaithersburg MD), to extend the nucleic acid sequence from an oUgonucleotide primer annealed to the DNA template of interest.
  • SEQUENASE DNA polymerase U.S. Biochemical Co ⁇ oration, Cleveland OH
  • Taq polymerase PE Biosystems, Foster City CA
  • thermostable T7 polymerase Amersham Pharmacia Biotech, Inc. (Amersham Pharma
  • Chain termination reaction products may be electrophoresed on urea-polyacrylamide gels and detected either by autoradiography (for radioisotope-labeled nucleotides) or by fluorescence (for fluorophore-labeled nucleotides).
  • Automated methods for mechanized reaction preparation, sequencing, and analysis using fluorescence detection methods have been developed.
  • Machines used to prepare cDNA for sequencing can include the MICROLAB 2200 liquid transfer system (Hamilton Company (Hamilton), Reno NV), Peltier thermal cycler (PTC200; MJ Research, Inc. (MJ Research), Watertown MA), and ABI CATALYST 800 thermal cycler (PE Biosystems). Sequencing can be carried out using, for example, the ABI 373 or 377 (PE Biosystems) or MEGABACE 1000 (Molecular Dynamics, Inc.
  • nucleotide sequences of the Sequence Listing have been prepared by current, state-of-the- art, automated methods and, as such, may contain occasional sequencing errors or unidentified nucleotides. Such unidentified nucleotides are designated by an N. These infrequent unidentified bases do not represent a hindrance to practicing the invention for those skilled in the art.
  • Several methods employing standard recombinant techniques may be used to correct errors and complete the missing sequence information. (See, e.g., those described in Ausubel, F.M. et al. (1997) Short Protocols in 5 Molecular Biology, John Wiley & Sons, New York NY; and Sambrook, J. et al. (1989) Molecular Cloning, A Laboratory Manual, Cold Spring Harbor Press, Plainview NY.)
  • Human polynucleotide sequences may be assembled using programs or algorithms well known 0 in the art. Sequences to be assembled are related, wholly or in part, and may be derived from a single or many different transcripts. Asembly of the sequences can be performed using such programs as PHRAP (Phils Revised Asembly Program) and the GEL VIEW fragment assembly system (GCG), or other methods known in the art.
  • PHRAP Phils Revised Asembly Program
  • GCG GEL VIEW fragment assembly system
  • cDNA sequences are used as "component" sequences that are assembled into 5 "template” or “consensus” sequences as follows. Sequence chromatograms are processed, verified, and quaUty scores are obtained using PHRED. Raw sequences are edited using an editing pathway known as Block 1 (See, e.g., theLIFESEQ Asembled User Guide, Incyte Genomics, Palo Ato, CA). A series of BLAST comparisons is performed and low-information segments and repetitive elements (e.g., dinucleotide repeats, Au repeats, etc.) are replaced by "n's", or masked, to prevent spurious matches.
  • Block 1 See, e.g., theLIFESEQ Asembled User Guide, Incyte Genomics, Palo Ato, CA).
  • a series of BLAST comparisons is performed and low-information segments and repetitive elements (e.g., dinucleotide repeats, Au repeats, etc.) are replaced by "n's
  • RNA sequences are also removed.
  • the processed sequences are then loaded into a relational database management system (RDMS) which assigns edited sequences to existing templates, if available.
  • RDMS relational database management system
  • a process is initiated which modifies existing templates or creates new templates from works in progress (i.e., nonfinal assembled sequences) containing queued sequences or the sequences themselves.
  • the templates can be merged into bins. If multiple templates exist in one bin, the bin can be spUt and the templates reannotated.
  • bins are "clone joined" based upon clone information. Clone joining occurs when the 5' sequence of one clone is present in one bin and the 3' sequence from the same clone is present in a different bin, indicating that the two bins o should be merged into a single bin. Only bins which share at least two different clones are merged.
  • a resultant template sequence may contain either a partial or a full length open reading frame, or all or part of a genetic regulatory element. This variation is due in part to the fact that the full length cDNA of many genes are several hundred, and sometimes several thousand, bases in length. With current technology, cDNA comprising the coding regions of large genes cannot be cloned because of vector Umitations, incomplete reverse transcription of the mRNA, or incomplete "second strand" synthesis. Template sequences may be extended to include additional contiguous sequences derived from the parent RNA transcript using a variety of methods known to those of skill in the art. Extension may thus be used to achieve the full length coding sequence of a gene.
  • the cDNA sequences are analyzed using a variety of programs and algorithms which are well known in the art. (See, e.g., Ausubel, 1997, supra. Chapter 7.7; Meyers, R.A. (Ed.) (1995) Molecular Biology and Biotechnology, Wiley VCH, New York NY, pp. 856-853; and Table 6.) These analyses o comprise both reading frame determinations, e.g. , based on triplet codon periodicity for particular organisms (Fickett, J.W. (1982) Nucleic Acids Res. 10:5303-5318); analyses of potential start and stop codons; and homology searches.
  • BLAST Basic Local 5 AUgnment Search Tool
  • BLAST is especially useful in determining exact matches and comparing two sequence fragments of arbitrary but equal lengths, whose aUgnment is locally maximal and for which the alignment score meets or exceeds a threshold or cutoff score set by the user (KarUn, S. et al. (1988) Proc. Natl. Acad. Sci.
  • GenBank e.g., GenBank, SwissProt, BLOCKS, PFAM and other databases may be searched for sequences containing regions of homology to a query dithp or DITHP of the present invention.
  • search tool e.g., o BLAST or HMM
  • GenBank e.g., GenBank, SwissProt, BLOCKS, PFAM and other databases may be searched for sequences containing regions of homology to a query dithp or DITHP of the present invention.
  • Protein hierarchies can be assigned to the putative encoded polypeptide based on, e.g., motif, BLAST, or biological analysis. Methods for assigning these hierarchies are described, for example, in o "Database System Employing Protein Function Hierarchies for Viewing Biomolecular Sequence Data,"
  • SEQ ID NO:l SEQ ID NO:2, SEQ ID NO:3, SEQ ID NO:4, SEQ ID NO:5, SEQ ID NO:6, SEQ ID NO:7, and SEQ ID NO:8 encode, for example, human enzyme molecules.
  • SEQ ID NO:9 encodes, for example, an extracellular information transmission molecule.
  • SEQ ID NO: 10 and SEQ ID NO: 11 encode, for example, receptor molecules.
  • SEQ ID NO:12, SEQ ID NO:13, SEQ ID NO:14, SEQ ID NO:15, SEQ ID NO:16, SEQ ID NO:17, and SEQ ID NO:18 encode, for example, intracellular signaling molecules.
  • SEQ ID NO:19, SEQ ID NO:20, SEQ ID NO:21, SEQ ID NO:22, SEQ ID NO:23, SEQ ID NO:24, SEQ ID NO:25, SEQ ID NO:26, SEQ ID NO:27, SEQ ID NO:28, SEQ ID NO:29, SEQ ID NO:30, SEQ ID NO:31, SEQ ID NO:32, and SEQ ID NO:33 encode, for example, ttanscription factor molecules.
  • SEQ ID NO:34 encodes, for example, a protein modification and maintenance molecule.
  • SEQ ID NO:35 and SEQ ID NO:36 encode, for example, nucleic acid synthesis and modification molecules.
  • SEQ ID NO:37 encodes, for example, an antigen recognition molecule.
  • SEQ ID NO:38 and SEQ ID NO:39 encode, for example, secreted/extracellular matrix molecules.
  • SEQ ID NO:40, SEQ ID NO:41, SEQ ID NO:42, SEQ ID NO:43, SEQ ID NO:44, and SEQ ID NO:45 encode, for example, cytoskeletal molecules.
  • SEQ ID NO:46, SEQ ID NO:47, and SEQ ID NO:48 encode, for example, cell membrane molecules.
  • SEQ ID NO:49, SEQ ID NO:50, SEQ ID NO:51, SEQ ID NO:52, and SEQ ID NO:53 encode, for example, ribosomal molecules.
  • SEQ ID NO:54, SEQ ID NO:55, SEQ ID NO:56, SEQ ID NO:57, SEQ ID NO:58, SEQ ID NO:59, SEQ ID NO:60, SEQ ID NO:61, SEQ ID NO:62, and SEQ ID NO:63 encode, for example, organelle associated molecules.
  • SEQ ID NO:64, SEQ ID NO:65, SEQ ID NO:66, SEQ ID NO:67, and SEQ ID NO:68 encode, for example, biochemical pathway molecules.
  • SEQ ID NO:69, SEQ ID NO:70, and SEQ ID NO:71 encode, for example, molecules associated with growth and development.
  • the dithp of the present invention may be used for a variety of diagnostic and therapeutic pu ⁇ oses.
  • a dithp may be used to diagnose a particular condition, disease, or disorder associated with human molecules.
  • Such conditions, diseases, and disorders include, but are not Umited to, a cell proliferative disorder, such as actinic keratosis, arteriosclerosis, atherosclerosis, bursitis, cirrhosis, hepatitis, mixed connective tissue disease (MCTD), myelofibrosis, paroxysmal nocturnal hemoglobinuria, polycythemia vera, psoriasis, primary thrombocythemia, and cancers including adenocarcinoma, leukemia, lymphoma, melanoma, myeloma, sarcoma, teratocarcinoma, and, in particular, a cancer of the adrenal gland, bladder, bone, bone marrow, brain, breast, cervix,
  • the dithp can be used to detect the presence of, or to quantify the amount of, a dithp-related polynucleotide in a sample. This information is then compared to information obtained from appropriate reference samples, and a diagnosis is estabUshed. Aternatively, a polynucleotide complementary to a given dithp can inhibit or inactivate a therapeutically relevant gene related to the dithp.
  • the expression of dithp may be routinely assessed by hybridization-based methods to determine, for example, the tissue-specificity, disease-specificity, or developmental stage-specificity of dithp expression.
  • the level of expression of dithp may be compared among different cell types or tissues, among diseased and normal cell types or tissues, among cell types or tissues at different developmental stages, or among cell types or tissues undergoing various treatments. This type of analysis is useful, for example, to assess the relative levels of dithp expression in fully or partially differentiated cells or tissues, to determine if changes in dithp expression levels are correlated with the development or progression of specific disease states, and to assess the response of a cell or tissue to a specific therapy, for example, in pharmacological or toxicological studies.
  • Methods for the analysis of dithp expression are based on hybridization and ampUfication technologies and include membrane- based procedures such as northern blot analysis, high-throughput procedures that utitize, for example, microarrays, and PCR-based procedures.
  • the dithp, their fragments, or complementary sequences may be used to identify the presence of and or to determine the degree of similarity between two (or more) nucleic acid sequences.
  • the dithp may be hybridized to naturally occurring or recombinant nucleic acid sequences under appropriately selected temperatures and salt concentrations.
  • Hybridization with a probe based on the nucleic acid sequence of at least one of the dithp allows for the detection of nucleic acid sequences, including genomic sequences, which are identical or related to the dithp of the Sequence Listing.
  • Probes may be selected from non-conserved or unique regions of at least one of the polynucleotides of SEQ ID NO:l- 71 and tested for their abiUty to identify or ampUfy the target nucleic acid sequence using standard protocols.
  • Polynucleotide sequences that are capable of hybridizing, in particular, to those shown in SEQ ID NO: 1-71 and fragments thereof, can be identified using various conditions of stringency. (See, e.g. , Wahl, G.M. and S.L. Berger (1987) Methods Enzymol. 152:399-407; Kimmel, AR. (1987) Methods Enzymol. 152:507-511.) Hybridization conditions are discussed in "Definitions.”
  • a probe for use in Southern or northern hybridization may be derived from a fragment of a dithp sequence, or its complement, that is up to several hundred nucleotides in length and is either single-stranded or double-stranded. Such probes may be hybridized in solution to biological materials such as plasmids, bacterial, yeast, or human artificial chromosomes, cleared or sectioned tissues, or to artificial substrates containing dithp. Microarrays are particularly suitable for identifying the presence of and detecting the level of expression for multiple genes of interest by examining gene expression 5 correlated with, e.g., various stages of development, treatment with a drug or compound, or disease progression.
  • An array analogous to a dot or slot blot may be used to arrange and tink polynucleotides to the surface of a substrate using one or more of the following: mechanical (vacuum), chemical, thermal, or UV bonding procedures.
  • Such an array may contain any number of dithp and may be produced by hand or by using available devices, materials, and machines.
  • Microarrays may be prepared, used, and analyzed using methods known in the art. (See, e.g.,
  • 5 Probes may be labeled by either PCR or enzymatic techniques using a variety of commercially available reporter molecules.
  • commercial kits are available for radioactive and chemiluminescent labeling (Amersham Pharmacia Biotech) and for alkaline phosphatase labeUng (Life Technologies).
  • dithp may be cloned into commercially available vectors for the production of RNA probes.
  • Such probes may be transcribed in the presence of at least one labeled o nucleotide (e.g. , 32 P-ATP, Amersham Pharmacia Biotech).
  • polynucleotides of SEQ ID NO: 1-71 or suitable fragments thereof can be used to isolate full length cDNA sequences utihzing hybridization and/or ampUfication procedures well known in the art, e.g., cDNA Ubrary screening, PCR ampUfication, etc.
  • the molecular cloning of such full length cDNA sequences may employ the method of cDNA library screening with probes using the 5 hybridization, stringency, washing, and probing strategies described above and in Ausubel, supra.
  • a genetic tinkage map traces parts of chromosomes that are inherited in the same pattern as the conditioa Statistics Unk the inheritance of particular conditions to particular regions of chromosomes, as defined by RFLP or other markers. 0 (See, for example, Lander, E. S. and Botstein, D. (1986) Proc. Natl. Acad. Sci. USA 83:7353-7357.) Occasionally, genetic markers and their locations are known from previous studies. More often, however, the markers are simply stretches of DNA that differ among individuals. Examples of genetic tinkage maps can be found in various scientific journals or at the Onhne Mendelian Inheritance in Man (OMIM) World Wide Web site.
  • OMIM Onhne Mendelian Inheritance in Man
  • dithp sequences may be used to generate hybridization probes useful in chromosomal mapping of naturally occurring genomic sequences. Either coding or noncoding sequences of dithp may be used, and in some instances, noncoding sequences may be preferable over coding sequences. For example, conservation of a dithp coding sequence among members of a multi-gene family may potentially cause undesired cross hybridization during o chromosomal mapping.
  • sequences may be mapped to a particular chromosome, to a specific region of a chromosome, or to artificial chromosome constructions, e.g., human artificial chromosomes (HACs), yeast artificial chromosomes (YACs), bacterial artificial chromosomes (BACs), bacterial PI constructions, or single chromosome cDNA Ubraries.
  • HACs human artificial chromosomes
  • YACs yeast artificial chromosomes
  • BACs bacterial artificial chromosomes
  • PI constructions or single chromosome cDNA Ubraries.
  • Fluorescent in situ hybridization may be correlated with other physical chromosome mapping techniques and genetic map data. (See, e.g., Meyers, supra, pp. 965-968.) Correlation between the location of dithp on a physical chromosomal map and a specific disorder, or a predisposition to a specific disorder, may help define the region of DNA associated with that disorder. o The dithp sequences may also be used to detect polymo ⁇ hisms that are genetically Unked to the inheritance of a particular condition, disease, or disorder.
  • In situ hybridization of chromosomal preparations and genetic mapping techniques may be used for extending existing genetic maps. Often the placement of a gene on the chromosome of another mammalian species, such as mouse, may reveal associated markers even if the number or arm of the corresponding human chromosome is not known. These new marker sequences can be mapped to human chromosomes and may provide valuable information to investigators searching for disease genes using positional cloning or other gene discovery techniques.
  • any sequences mapping to that area may represent associated or regulatory genes for further investigation.
  • the nucleotide sequences of the subject invention may also be used to detect differences in chromosomal architecture due to translocation, inversion, etc., among normal, carrier, or affected individuals.
  • a disease-associated gene is mapped to a chromosomal region, the gene must be cloned in order to identify mutations or other alterations (e.g., ttanslocations or inversions) that may be correlated with disease.
  • This process requires a physical map of the chromosomal region containing the disease- gene of interest along with associated markers. A physical map is necessary for determining the nucleotide sequence of and order of marker genes on a particular chromosomal region. Physical 5 mapping techniques are well known in the art and require the generation of overlapping sets of cloned DNA fragments from a particular organelle, chromosome, or genome. These clones are analyzed to reconstruct and catalog their order. Once the position of a marker is determined, the DNA from that region is obtained by consulting the catalog and selecting clones from that region. The gene of interest is located through positional cloning techniques using hybridization or similar methods.
  • the dithp of the present invention may be used to design probes useful in diagnostic assays. Such assays, well known to those skilled in the art, may be used to detect or confirm conditions, disorders, or diseases associated with abnormal levels of dithp expression. Labeled probes developed 5 from dithp sequences are added to a sample under hybridizing conditions of desired sttingency. In some instances, dithp, or fragments or oUgonucleotides derived from dithp, may be used as primers in ampUfication steps prior to hybridization. The amount of hybridization complex formed is quantified and compared with standards for that cell or tissue. If dithp expression varies significantly from the standard, the assay indicates the presence of the condition, disorder, or disease.
  • QuaUtative or o quantitative diagnostic methods may include northern, dot blot, or other membrane or dip-stick based technologies or multiple-sample format technologies such as PCR, enzyme-Unked immunosorbent assay (ELISA)-Uke, pin, or chip-based assays.
  • PCR enzyme-Unked immunosorbent assay
  • the probes described above may also be used to monitor the progress of conditions, disorders, or diseases associated with abnormal levels of dithp expression, or to evaluate the efficacy of a particular therapeutic treatment.
  • the candidate probe may be identified from the dithp that are specific to a given human tissue and have not been observed in GenBank or other genome databases. Such a probe may be used in animal studies, precUnical tests, cUnical trials, or in monitoring the treatment of an individual patient.
  • standard expression is estabUshed by methods well known in 5 the art for use as a basis of comparison, samples from patients affected by the disorder or disease are combined with the probe to evaluate any deviation from the standard profile, and a therapeutic agent is administered and effects are monitored to generate a treatment profile.
  • Efficacy is evaluated by determining whether the expression progresses toward or returns to the standard normal pattern. Treatment profiles may be generated over a period of several days or several months. Statistical o methods well known to those skilled in the art may be use to determine the significance of such therapeutic agents.
  • the polynucleotides are also useful for identifying individuals from minute biological samples, for example, by matching the RFLP pattern of a sample's DNA to that of an individual's DNA
  • the polynucleotides of the present invention can also be used to determine the actual base-by-base DNA 5 sequence of selected portions of an individual's genome. These sequences can be used to prepare PCR primers for ampUfying and isolating such selected DNA, which can then be sequenced. Using this technique, an individual can be identified through a unique set of DNA sequences. Once a unique ID database is established for an individual, positive identification of that individual can be made from extremely small tissue samples.
  • oUgonucleotide primers derived from the dithp of the invention may be used to detect single nucleotide polymo ⁇ hisms (SNPs).
  • SNPs are substitutions, insertions and deletions that are a frequent cause of inherited or acquired genetic disease in humans.
  • Methods of SNP detection include, but are not Umited to, single-stranded conformation polymorphism (SSCP) and fluorescent SSCP (fSSCP) methods.
  • SSCP single-stranded conformation polymorphism
  • fSSCP fluorescent SSCP
  • oligonucleotide primers derived from dithp are used to 5 ampUfy DNA using the polymerase chain reaction (PCR).
  • the DNA may be derived, for example, from diseased or normal tissue, biopsy samples, bodily fluids, and the like.
  • SNPs in the DNA cause differences in the secondary and tertiary structures of PCR products in single-stranded form, and these differences are detectable using gel electrophoresis in non-denaturing gels.
  • the oUgonucleotide primers are fluorescently labeled, which allows detection of the ampUmers in high- o throughput equipment such as DNA sequencing machines.
  • sequence database analysis methods termed in siUco SNP (isSNP) are capable of identifying polymo ⁇ hisms by comparing the sequences of individual overlapping DNA fragments which assemble into a common consensus sequence.
  • SNPs may be detected and characterized by mass spectrometry using, for example, the high throughput MASSARRAY system (Sequenom, Inc., San Diego CA).
  • DNA-based identification techniques are critical in forensic technology. DNA sequences taken from very small biological samples such as tissues, e.g., hair or skin, or body fluids, e.g., blood, saliva, 5 semen, etc., can be ampUfied using, e.g., PCR, to identify individuals. (See, e.g., Erlich, H. (1992) PCR Technology, Freeman and Co., New York, NY). Similarly, polynucleotides of the present invention can be used as polymo ⁇ hic markers.
  • reagents capable of identifying the source of a particular tissue.
  • Appropriate reagents can comprise, for example, DNA probes or primers prepared from the sequences 0 of the present invention that are specific for particular tissues. Panels of such reagents can identify tissue by species and/or by organ type. In a similar fashion, these reagents can be used to screen tissue cultures for contamination.
  • polynucleotides of the present invention can also be used as molecular weight markers on nucleic acid gels or Southern blots, as diagnostic probes for the presence of a specific mRNA in a 5 particular cell type, in the creation of subtracted cDNA Ubraries which aid in the discovery of novel polynucleotides, in selection and synthesis of oUgomers for attachment to an array or other support, and as an antigen to elicit an immune response.
  • the dithp of the invention or their mammaUan homologs may be "knocked out" in an animal model system using homologous recombination in embryonic stem (ES) cells.
  • ES embryonic stem
  • Such techniques are well known in the art and are useful for the generation of animal models of human disease.
  • mouse ES cells such as the mouse 129/SvJ cell Une, are derived from the early mouse embryo and grown in culture.
  • the ES 5 cells are transformed with a vector containing the gene of interest disrupted by a marker gene, e.g., the neomycin phosphotransferase gene (neo; Capecchi, M.R. (1989) Science 244:1288-1292).
  • the vector integrates into the corresponding region of the host genome by homologous recombination. Aternatively, homologous recombination takes place using the Cre-loxP system to knockout a gene of interest in a tissue- or developmental stage-specific manner (Marth, J.D. (1996) Clin. Invest. 97:1999- o 2002; Wagner, K.U. et al. (1997) Nucleic Acids Res. 25 :4323-4330).
  • Transformed ES cells are identified and microinjected into mouse cell blastocysts such as those from the C57BL/6 mouse strain.
  • the blastocysts are surgically transferred to pseudopregnant dams, and the resulting chimeric progeny are genotyped and bred to produce heterozygous or homozygous strains.
  • Transgenic animals thus generated may be tested with potential therapeutic or toxic agents.
  • the dithp of the invention may also be manipulated in vitro in ES cells derived from human blastocysts.
  • Human ES cells have the potential to differentiate into at least eight separate cell Uneages including endoderm, mesoderm, and ectodermal cell types. These cell Uneages differentiate into, for example, neural cells, hematopoietic Uneages, and cardiomyocytes (Thomson, J.A. et al. (1998) Science 5 282:1145-1147).
  • the dithp of the invention can also be used to create "knockin" humanized animals (pigs) or transgenic animals (mice or rats) to model human disease.
  • knockin technology a region of dithp is injected into animal ES cells, and the injected sequence integrates into the animal cell genome.
  • Transformed cells are injected into blastulae, and the blastulae are implanted as described above.
  • Transgenic progeny or inbred lines are studied and treated with potential pharmaceutical agents to obtain information on treatment of a human disease.
  • a mammal inbred to overexpress dithp resulting, e.g., in the secretion of DITHP in its milk, may also serve as a convenient source of that protein (Janne, J. et al. (1998) Biotechnol. Ainu. Rev. 4:55-74).
  • DITHP encoded by polynucleotides of the present invention may be used to screen for molecules that bind to or are bound by the encoded polypeptides.
  • the binding of the polypeptide and the molecule may activate (agonist), increase, inhibit (antagonist), or decrease activity of the polypeptide or the bound molecule.
  • Examples of such molecules include antibodies, oUgonucleotides, 0 proteins (e.g., receptors), or small molecules.
  • the molecule is closely related to the natural Ugand of the polypeptide, e.g., a Ugand or fragment thereof, a natural substrate, or a structural or functional mimetic.
  • the molecule can be closely related to the natural receptor to which the polypeptide binds, or to at least a fragment of the receptor, 5 e.g., the active site. In either case, the molecule can be rationally designed using known techniques.
  • the screening for these molecules involves producing appropriate cells which express the polypeptide, either as a secreted protein or on the cell membrane.
  • Preferred cells include cells from mammals, yeast, Drosophila, or E. coli. Cells expressing the polypeptide or cell membrane fractions which contain the expressed polypeptide are then contacted with a test compound and binding, o stimulation, or inhibition of activity of either the polypeptide or the molecule is analyzed.
  • Ai assay may simply test binding of a candidate compound to the polypeptide, wherein binding is detected by a fluorophore, radioisotope, enzyme conjugate, or other detectable label. Aternatively, the assay may assess binding in the presence of a labeled competitor. Additionally, the assay can be carried out using cell-free preparations, polypeptide/molecule affixed to a soUd support, chemical libraries, or natural product mixtures. The assay may also simply comprise the steps of mixing a candidate compound with a solution containing a polypeptide, measuring polypeptide/molecule activity or binding, and comparing the polypeptide/molecule activity or binding to 5 a standard.
  • an ELISA assay using, e.g., a monoclonal or polyclonal antibody can measure polypeptide level in a sample.
  • the antibody can measure polypeptide level by either binding, directly or indirectly, to the polypeptide or by competing with the polypeptide for a substrate.
  • Al of the above assays can be used in a diagnostic or prognostic context.
  • the molecules o discovered using these assays can be used to treat disease or to bring about a particular result in a patient (e.g., blood vessel growth) by activating or inhibiting the polypeptide/molecule.
  • the assays can discover agents which may inhibit or enhance the production of the polypeptide from suitably manipulated cells or tissues.
  • Aiother embodiment relates to the use of dithp to develop a transcript image of a tissue or cell type.
  • a transcript image represents the global pattern of gene expression by a particular tissue or cell type. Global gene expression patterns are analyzed by quantifying the number of expressed genes and their relative abundance under given conditions and at a given time. (See Seilhamer et al., o "Comparative Gene Transcript Analysis," U.S. Patent Number 5,840,484, expressly inco ⁇ orated by reference herein.)
  • a transcript image may be generated by hybridizing the polynucleotides of the present invention or their complements to the totaUty of transcripts or reverse transcripts of a particular tissue or cell type.
  • the hybridization takes place in high-throughput format, wherein the polynucleotides of the present invention or their complements comprise a subset of a 5 plurality of elements on a microarray.
  • the resultant transcript image would provide a profile of gene activity pertaining to human molecules for diagnostics and therapeutics.
  • Transcript images which profile dithp expression may be generated using transcripts isolated from tissues, cell Unes, biopsies, or other biological samples.
  • the transcript image may thus reflect dithp expression in vivo, as in the case of a tissue or biopsy sample, or in vitro, as in the case of a cell 0 Une.
  • Transcript images which profile dithp expression may also be used in conjunction with in vitro model systems and preclinical evaluation of pharmaceuticals, as well as toxicological testing of industrial and naturally-occurring environmental compounds.
  • Al compounds induce characteristic gene expression patterns, frequently termed molecular finge ⁇ rints or toxicant signatures, which are indicative of mechanisms of action and toxicity (Nuwaysir, E. F. et al. (1999) Mol. Carcinog. 24:153- 159; Steiner, S. and Aiderson, N. L. (2000) Toxicol. Lett. 112-113:467-71, expressly inco ⁇ orated by reference herein). If a test compound has a signature similar to that of a compound with known toxicity, it is Ukely to share those toxic properties.
  • finge ⁇ rints or signatures are most useful and 5 refined when they contain expression information from a large number of genes and gene famiUes. Ideally, a genome-wide measurement of expression provides the highest quality signature. Even genes whose expression is not altered by any tested compounds are important as well, as the levels of expression of these genes are used to normatize the rest of the expression data. The normatization procedure is useful for comparison of expression data after treatment with different compounds. While 0 the assignment of gene function to elements of a toxicant signature aids in inte ⁇ retation of toxicity mechanisms, knowledge of gene function is not necessary for the statistical matching of signatures which leads to prediction of toxicity.
  • the toxicity of a test compound is assessed by treating a biological sample containing nucleic acids with the test compound.
  • Nucleic acids that are expressed in the treated biological sample are hybridized with one or more probes specific to the polynucleotides of the present invention, so that transcript levels corresponding to the polynucleotides of the present o invention may be quantified.
  • the transcript levels in the treated biological sample are compared with levels in an untreated biological sample. Differences in the transcript levels between the two samples are indicative of a toxic response caused by the test compound in the treated sample.
  • Aiother particular embodiment relates to the use of DITHP encoded by polynucleotides of the present invention to analyze the proteome of a tissue or cell type.
  • proteome refers to the 5 global pattern of protein expression in a particular tissue or cell type. Each protein component of a proteome can be subjected individually to further analysis. Proteome expression patterns, or profiles, are analyzed by quantifying the number of expressed proteins and their relative abundance under given conditions and at a given time. A profile of a cell's proteome may thus be generated by separating and analyzing the polypeptides of a particular tissue or cell type.
  • the separation is o achieved using two-dimensional gel electtophoresis, in which proteins from a sample are separated by isoelectric focusing in the first dimension, and then according to molecular weight by sodium dodecyl sulfate slab gel electrophoresis in the second dimension (Steiner and Anderson, supra).
  • the proteins are visuahzed in the gel as discrete and uniquely positioned spots, typically by staining the gel with an agent such as Coomassie Blue or silver or fluorescent stains.
  • the optical density of each protein spot is generally proportional to the level of the protein in the sample.
  • the optical densities of equivalently positioned protein spots from different samples are compared to identify any changes in protein spot density related to the treatment.
  • the proteins in the spots are partially sequenced using, for example, standard methods employing chemical or enzymatic cleavage followed by mass spectiOmetry.
  • the identity of the protein in a spot may be determined by comparing its partial sequence, preferably of at least 5 contiguous amino acid residues, to the polypeptide sequences of the present invention. In some cases, further sequence data may be obtained for definitive protein identification.
  • a proteomic profile may also be generated using antibodies specific for DITHP to quantify the levels of DITHP expression.
  • the antibodies are used as elements on a microarray, and protein expression levels are quantified by exposing the microarray to the sample and detecting the levels of protein bound to each array element (Lueking, A. et al. (1999) Anal. Biochem. 270:103-11; Mendoze, L. G. et al. (1999) Biotechniques 27:778-88). Detection may be performed by a variety of methods known in the art, for example, by reacting the proteins in the sample with a thiol- or amino- reactive fluorescent compound and detecting the amount of fluorescence bound at each array element.
  • Toxicant signatures at the proteome level are also useful for toxicological screening, and should be analyzed in parallel with toxicant signatures at the transcript level.
  • There is a poor correlation between transcript and protein abundances for some proteins in some tissues (Anderson, N. L. and Seilhamer, J. (1997) Electrophoresis 18:533-537), so proteome toxicant signatures may be useful in the analysis of compounds which do not significantly affect the transcript image, but which alter the proteomic profile.
  • the analysis of transcripts in body fluids is difficult, due to rapid degradation of mRNA, so proteomic profiling may be more reUable and informative in such cases.
  • the toxicity of a test compound is assessed by treating a biological sample containing proteins with the test compound.
  • Proteins that are expressed in the treated biological sample are separated so that the amount of each protein can be quantified.
  • the amount of each protein is compared to the amount of the corresponding protein in an untreated biological sample. A difference in the amount of protein between the two samples is indicative of a toxic response to the test compound in the treated sample.
  • Individual proteins are identified by sequencing the amino acid residues of the individual proteins and comparing these partial sequences to the DITHP encoded by polynucleotides of the present invention.
  • the toxicity of a test compound is assessed by treating a biological sample containing proteins with the test compound. Proteins from the biological sample are incubated with antibodies specific to the DITHP encoded by polynucleotides of the present invention. The amount of protein recognized by the antibodies is quantified. The amount of protein in the treated biological sample is compared with the amount in an untreated biological sample. A difference in the amount of protein between the two samples is indicative of a toxic response to the test compound in the treated sample.
  • Transcript images may be used to profile dithp expression in distinct tissue types. This process 5 can be used to determine human molecule activity in a particular tissue type relative to this activity in a different tissue type. Transcript images may be used to generate a profile of dithp expression characteristic of diseased tissue. Transcript images of tissues before and after treatment may be used for diagnostic pu ⁇ oses, to monitor the progression of disease, and to monitor the efficacy of drug treatments for diseases which affect the activity of human molecules. o Transcript images of cell Unes can be used to assess human molecule activity and/or to identify cell lines that lack or misregulate this activity. Such cell Unes may then be treated with pharmaceutical agents, and a transcript image following treatment may indicate the efficacy of these agents in restoring desired levels of this activity. A similar approach may be used to assess the toxicity of pharmaceutical agents as reflected by undesirable changes in human molecule activity. Candidate pharmaceutical 5 agents may be evaluated by comparing their associated transcript images with those of pharmaceutical agents of known effectiveness.
  • the polynucleotides of the present invention are useful in antisense technology.
  • Antisense o technology or therapy reUes on the modulation of expression of a target protein through the specific binding of an antisense sequence to a target sequence encoding the target protein or directing its expression.
  • Agrawal, S., ed. 1996 Antisense Therapeutics, Humana Press Inc., Totawa NJ; Aa a, A. et al. (1997) Pharmacol. Res. 36(3):171-178; Crooke, S.T. (1997) Adv. Pharmacol. 40:1-49; Sharma, H.W. and R.
  • An antisense sequence is a polynucleotide sequence capable of specifically hybridizing to at least a portion of the target sequence. Antisense sequences bind to cellular mRNA and/or genomic DNA, affecting translation and/or transcription. Antisense sequences can be DNA, RNA, or nucleic acid mimics and analogs. (See, e.g., Rossi, J.J. et al. (1991) Antisense Res. Dev. l(3):285-288; Lee, R. et al.
  • the binding which results in modulation of expression occurs through hybridization or binding of complementary base pairs.
  • Antisense sequences can also bind to DNA duplexes through specific interactions in the major groove of the double helix.
  • the polynucleotides of the present invention and fragments thereof can be used as antisense sequences to modify the expression of the polypeptide encoded by dithp.
  • antisense sequences can be produced ex vivo, such as by using any of the ABI nucleic acid synthesizer series (PE Biosystems) or other automated systems known in the art. Antisense sequences can also be produced biologically, such as by transforming an appropriate host cell with an expression vector containing the sequence of interest. (See, e.g., Agrawal, supra.)
  • Antisense sequences can be delivered inttacellularly in the form of an expression plasmid which, upon transcription, produces a sequence complementary to at least a portion of the cellular sequence encoding the target protein.
  • Aitisense sequences can also be introduced inttacellularly through the use of viral vectors, such as retrovirus and adeno-associated virus vectors.
  • viral vectors such as retrovirus and adeno-associated virus vectors.
  • retrovirus vectors See, e.g., Miller, A.D. (1990) Blood 76:271; Ausubel, F.M. et al. (1995) Current Protocols in Molecular Biology. John Wiley & Sons, New York NY; Uckert, W. and W. Walther (1994) Pharmacol. Ther. 63(3):323-347.
  • Other gene delivery mechanisms include liposome-derived systems, artificial viral envelopes, and other systems known in the art. (See, e.g., Rossi, J.J. (1995) Br. Med.
  • the nucleotide sequences encoding DITHP or fragments thereof may be inserted into an appropriate expression vector, i.e., a vector which contains the necessary elements for transcriptional and translational control of the inserted coding sequence in a suitable host.
  • an appropriate expression vector i.e., a vector which contains the necessary elements for transcriptional and translational control of the inserted coding sequence in a suitable host.
  • Methods which are well known to those skilled in the art may be used to construct expression vectors containing sequences encoding DITHP and appropriate transcriptional and translational control elements. These methods include in vitro recombinant DNA techniques, synthetic techniques, and in vivo genetic recombination. (See, e.g., Sambrook, supra. Chapters 4, 8, 16, and 17; and Ausubel, supra, Chapters 9, 10, 13, and 16.)
  • a variety of expression vector/host systems may be utiUzed to contain and express sequences encoding DITHP. These include, but are not Umited to, microorganisms such as bacteria ttansformed with recombinant bacteriophage, plasmid, or cosmid DNA expression vectors; yeast ttansformed with yeast expression vectors; insect cell systems infected with viral expression vectors (e.g., baculovirus); plant cell systems transformed with viral expression vectors (e.g., cauliflower mosaic virus, CaMV, or tobacco mosaic virus, TMV) or with bacterial expression vectors (e.g., Ti or pBR322 plasmids); or animal (mammalian) cell systems.
  • microorganisms such as bacteria ttansformed with recombinant bacteriophage, plasmid, or cosmid DNA expression vectors
  • yeast ttansformed with yeast expression vectors e.g., baculovirus
  • Expression vectors derived from rettoviruses, adenoviruses, 0 or he ⁇ es or vaccinia viruses, or from various bacterial plasmids may be used for deUvery of nucleotide sequences to the targeted organ, tissue, or cell population.
  • the invention is not 5 Umited by the host cell employed.
  • sequences encoding DITHP can be ttansformed into cell Unes using expression vectors which may contain viral origins of repUcation and/or endogenous expression elements and a selectable marker gene on the same or on a separate vector. Any number of o selection systems may be used to recover transformed cell Unes. (See, e.g., Wigler, M. et al. (1977)
  • the dithp of the invention may be used for somatic or germUne gene therapy.
  • Gene therapy may be performed to (i) correct a genetic deficiency (e.g., in the cases of severe combined immunodeficiency (SCID)-Xl disease characterized by X-Unked inheritance (Cavazzana-Calvo, M. et o al. (2000) Science 288 :669-672), severe combined immunodeficiency syndrome associated with an inherited adenosine deaminase (ADA) deficiency (Blaese, R.M. et al. (1995) Science 270:475-480; Bordignon, C et al.
  • SCID severe combined immunodeficiency
  • ADA adenosine deaminase
  • hepatitis B or C virus HBV, HCV
  • fungal parasites such as Candida albicans and Paracoccidioides brasihensis
  • protozoan parasites such as Plasmodium falciparum and Trvpanosoma cruzi.
  • the expression of dithp from an appropriate population of o transduced cells may alleviate the clinical manifestations caused by the genetic deficiency.
  • diseases or disorders caused by deficiencies in dithp are treated by constructing mammaUan expression vectors comprising dithp and introducing these vectors by mechanical means into dithp-deficient cells.
  • Mechanical transfer technologies for use with cells in vivo or ex vitro include (i) direct DNA microinjection into individual cells, (u) balUstic gold 5 particle delivery, (ui) Uposome-mediated transfection, (iv) receptor-mediated gene transfer, and (v) the use of DNA transposons (Morgan, R.A. and Anderson, W.F. (1993) Annu. Rev. Biochem. 62:191-217; Ivies, Z. (1997) Cell 91:501-510; Boulay, J-L. and R ⁇ cipon, H. (1998) Curr. Opin. Biotechnol. 9:445- 450).
  • Expression vectors that may be effective for the expression of dithp include, but are not Umited o to, the PCDNA 3.1, EPITAG, PRCCMV2, PREP, PVAX vectors (Invitrogen, Carlsbad C A),
  • the dithp of the invention may be expressed using (i) a constitutively active promoter, (e.g., from cytomegalovirus (CMV), Rous sarcoma virus (RSV), SV40 virus, thymidine kinase (TK), or ⁇ -actin genes), (n) an inducible promoter 5 (e.g., the tetracycUne-regulated promoter (Gossen, M. and Bujard, H.
  • a constitutively active promoter e.g., from cytomegalovirus (CMV), Rous sarcoma virus (RSV), SV40 virus, thymidine kinase (TK), or ⁇ -actin genes
  • an inducible promoter 5 e.g., the tetracycUne-regulated promoter (Gossen, M. and Bujard, H.
  • Uposome transformation kits e.g., the PERFECT LIPID TRANSFECTION KIT, available from Invitrogen
  • Uposome transformation allows one with ordinary skill in the art to deliver polynucleotides to target cells in culture and require minimal effort to optimize experimental parameters.
  • transformation is performed using the calcium phosphate method (Graham, F.L. andEb, AJ. (1973) Virology 52:456-467), or by electroporation (Neumann, E. et al. (1982) EMBO J. 1:841-845).
  • the introduction of DNA to primary cells requires modification of these standardized mammaUan transfection protocols.
  • diseases or disorders caused by genetic defects with respect to dithp expression are treated by constructing a rettovirus vector consisting of (i) dithp under the control of an independent promoter or the rettovirus long terminal repeat (LTR) promoter, (U) appropriate RNA packaging signals, and (iti) a Rev-responsive element (RRE) along with additional rettovirus cis-acting RNA sequences and coding sequences required for efficient vector propagation.
  • a rettovirus vector consisting of (i) dithp under the control of an independent promoter or the rettovirus long terminal repeat (LTR) promoter, (U) appropriate RNA packaging signals, and (iti) a Rev-responsive element (RRE) along with additional rettovirus cis-acting RNA sequences and coding sequences required for efficient vector propagation.
  • RRE Rev-responsive element
  • the vector is propagated in an appropriate vector producing cell Une (VPCL) that expresses an envelope gene with a tropism for receptors on the target cells or a promiscuous envelope protein such as VSVg (Amentano, D. et al. (1987) J. Virol. 61:1647-1650; Bender, M.A. et al. (1987) 5 J. Virol. 61 :1639-1646; Adam, M.A and Miller, AD. (1988) J. Virol. 62:3802-3806; Dull, T. et al. (1998) J. Virol. 72:8463-8471; Zufferey, R.
  • VPCL vector producing cell Une
  • U.S. Patent Number 5,910,434 to Rigg discloses a method for obtaining retrovirus packaging cell Unes and is hereby inco ⁇ orated by reference. Propagation of retrovirus vectors, transduction of a population of o cells (e.g., CD4 + T-cells), and the return of ttansduced cells to a patient are procedures well known to persons skilled in the art of gene therapy and have been well documented (Ranga, U. et al. (1997) J. Virol. 71:7020-7029; Bauer, G. et al.
  • an adenovirus-based gene therapy delivery system is used to deUver dithp to cells which have one or more genetic abnormaUties with respect to the expression of dithp.
  • the construction and packaging of adenovirus-based vectors are well known to those with ordinary skill in the art.
  • RepUcation defective adenovirus vectors have proven to be versatile for importing genes encoding immunoregulatory proteins into intact islets in the pancreas (Csete, M.E. et al. (1995) o Transplantation 27:263-268). Potentially useful adenoviral vectors are described in U.S. Patent
  • Adenovirus vectors for gene therapy hereby inco ⁇ orated by reference.
  • adenoviral vectors see also Antinozzi, P.A et al. (1999) Ainu. Rev. Nutr. 19:511-544 and Verma, I.M. and Somia, N. (1997) Nature 18:389:239-242, both inco ⁇ orated by reference herein.
  • a herpes-based, gene therapy deUvery system is used to deliver dithp to target cells which have one or more genetic abnormatities with respect to the expression of dithp.
  • HSV he ⁇ es simplex virus
  • the construction and packaging of 5 he ⁇ es-based vectors are well known to those with ordinary skill in the art.
  • a repUcation-competent he ⁇ es simplex virus (HSV) type 1 -based vector has been used to deUver a reporter gene to the eyes of primates (Liu, X. et al. (1999) Exp. Eye Res.l69:385-395).
  • the construction of a HSV-1 virus vector has also been disclosed in detail in U.S.
  • Patent Number 5,804,413 to DeLuca (“He ⁇ es simplex virus strains for gene transfer"), which is hereby inco ⁇ orated by reference.
  • U.S. Patent Number 5,804,413 o teaches the use of recombinant HSV d92 which consists of a genome containing at least one exogenous gene to be transferred to a cell under the control of the appropriate promoter for pu ⁇ oses including human gene therapy.
  • Aso taught by this patent are the construction and use of recombinant HSV strains deleted for ICP4, ICP27 and ICP22.
  • an alphavirus (positive, single-stranded RNA virus) vector is used to o deUver dithp to target cells.
  • SFV SemUki Forest Virus
  • alphavirus RNA repUcation a subgenomic RNA is generated that normally encodes the viral capsid proteins.
  • This subgenomic RNA repUcates to higher levels than the full-length genomic RNA, resulting in the overproduction of capsid proteins 5 relative to the viral proteins with enzymatic activity (e.g., protease and polymerase).
  • enzymatic activity e.g., protease and polymerase.
  • inserting dithp into the alphavirus genome in place of the capsid-coding region results in the production of a large number of dithp RNA and the synthesis of high levels of DITHP in vector transduced cells.
  • alphavirus infection is typically associated with cell lysis within a few days
  • the abiUty to estabUsh a persistent infection in hamster normal kidney cells (BHK-21) with a variant of Sindbis virus o (SIN) indicates that the lytic repUcation of alphaviruses can be altered to suit the needs of the gene therapy application (Dryga, S.A. et al. (1997) Virology 228:74-83).
  • the wide host range of alphaviruses will allow the introduction of dithp into a variety of cell types.
  • the specific transduction of a subset of cells in a population may require the sorting of cells prior to transduction.
  • the methods of manipulating infectious cDNA clones of alphaviruses, performing alphavirus cDNA and RNA ttansfections, and performing alphavirus infections, are well known to those with ordinary skill in the art.
  • Anti-DITHP antibodies may be used to analyze protein expression levels. Such antibodies include, but are not limited to, polyclonal, monoclonal, chimeric, single chain, and Fab fragments. For descriptions of and protocols of antibody technologies, see, e.g., Pound J.D. (1998) Immunochemical Protocols, Humana Press, Totowa, NJ.
  • amino acid sequence encoded by the dithp of the Sequence Listing may be analyzed by o appropriate software (e.g. , LASERGENE NAVIGATOR software, DNASTAR) to determine regions of high immunogenicity.
  • the optimal sequences for immunization are selected from the C-terminus, the N-terminus, and those intervening, hydrophiUc regions of the polypeptide which are likely to be exposed to the external environment when the polypeptide is in its natural conformation. Analysis used to select appropriate epitopes is also described by Ausubel (1997, supra. Chapter 11.7). Peptides used for 5 antibody induction do not need to have biological activity; however, they must be antigenic.
  • Peptides used to induce specific antibodies may have an amino acid sequence consisting of at five amino acids, preferably at least 10 amino acids, and most preferably 15 amino acids.
  • a peptide which mimics an antigenic fragment of the natural polypeptide may be fused with another protein such as keyhole Umpet cyanin (KLH; Sigma, St. Louis MO) for antibody production.
  • KLH keyhole Umpet cyanin
  • a peptide encompassing an antigenic o region may be expressed from a dithp, synthesized as described above, or purified from human cells.
  • mice, goats, and rabbits may be immunized by injection with a peptide.
  • various adjuvants may be used to increase immunological response.
  • peptides about 15 residues in length may be synthesized using an ABI 431 A 5 peptide synthesizer (PE Biosystems) using fmoc-chemistry and coupled to KLH (Sigma) by reaction with M-maleimidobenzoyl-N-hydroxysuccinimide ester (Ausubel, 1995, supra).
  • Rabbits are immunized with the peptide-KLH complex in complete Freund's adjuvant.
  • the resulting antisera are tested for antipeptide activity by binding the peptide to plastic, blocking with 1 % bovine serum albumin (BSA), reacting with rabbit antisera, washing, and reacting with radioiodinated goat anti-rabbit IgG.
  • BSA bovine serum albumin
  • Antisera with antipeptide activity are tested for anti-DITHP activity using protocols well known in the art, including ELISA, radioimmunoassay (RIA), and immunoblotting.
  • isolated and purified peptide may be used to immunize mice (about 100 ⁇ g of peptide) or rabbits (about 1 mg of peptide). Subsequently, the peptide is radioiodinated and used to screen the immunized animals' B-lymphocytes for production of antipeptide antibodies. Positive cells are then used to produce hybridomas using standard techniques. About 20 mg of peptide is sufficient for labeUng and screening several thousand clones. Hybridomas of interest are detected by screening with radioiodinated peptide to identify those fusions producing peptide-specific monoclonal antibody.
  • wells of a multi-well plate (FAST, Becton-Dickinson, Palo Ato, CA) 5 are coated with affinity-purified, specific rabbit-anti-mouse (or suitable anti-species IgG) antibodies at 10 mg/ml.
  • the coated wells are blocked with 1 % BSA and washed and exposed to supematants from hybridomas. After incubation, the wells are exposed to radiolabeled peptide at 1 mg/ml.
  • Clones producing antibodies bind a quantity of labeled peptide that is detectable above background. Such clones are expanded and subjected to 2 cycles of cloning. Cloned hybridomas are o injected into pristane-tteated mice to produce ascites, and monoclonal antibody is purified from the ascitic fluid by affinity chromatography on protein A (Amersham Pharmacia Biotech). Several procedures for the production of monoclonal antibodies, including in vitro production, are described in Pound (supra). Monoclonal antibodies with antipeptide activity are tested for anti-DITHP activity using protocols well known in the art, including ELISA, RIA, and immunoblotting.
  • Aitibody fragments containing specific binding sites for an epitope may also be generated.
  • such fragments include, but are not Umited to, the F(ab')2 fragments produced by pepsin digestion of the antibody molecule, and the Fab fragments generated by reducing the disulfide bridges of the F(ab')2 fragments.
  • construction of Fab expression libraries in filamentous bacteriophage allows rapid and easy identification of monoclonal fragments with desired specificity 0 (Pound, supra. Chaps. 45-47).
  • Aitibodies generated against polypeptide encoded by dithp can be used to purify and characterize full-length DITHP protein and its activity, binding partners, etc.
  • Aiti-DITHP antibodies may be used in assays to quantify the amount of DITHP found in a 5 particular human cell. Such assays include methods utiUzing the antibody and a label to detect expression level under normal or disease conditions.
  • the peptides and antibodies of the invention may be used with or without modification or labeled by joining them, either covalently or noncovalently, with a reporter molecule.
  • Protocols for detecting and measuring protein expression using either polyclonal or monoclonal o antibodies are well known in the art. Examples include ELISA, RIA, and fluorescent activated cell sorting (FACS). Such immunoassays typically involve the formation of complexes between the DITHP and its specific antibody and the measurement of such complexes. These and other assays are described in Pound (supra). Without further elaboration, it is beUeved that one skilled in the art can, using the preceding description, utiUze the present invention to its fullest extent. The following preferred specific embodiments are, therefore, to be construed as merely illusttative, and not Umitative of the remainder of the disclosure in any way whatsoever.
  • 60/167,945, U.S. Ser. No. 60/167,520, U.S. Ser. No. 60/168,468, U.S. Ser. No. 60/168,599, U.S. Ser. No. 60/167,410, U.S. Ser. No. 60/168,265, U.S. Ser. No. 60/168,429, U.S. Ser. No. 60/168,432, U.S. Ser. No. 60/167,521, U.S. Ser. No. 60/168,857, U.S. Ser. No. 60/168,197, U.S. Ser. No. 60/168,611, and U.S. Ser. No. 60/168,613 are hereby expressly inco ⁇ orated by reference.
  • RNA was purchased from CLONTECH Laboratories, Inc. (Palo Ato CA) or isolated from various tissues. Some tissues were homogenized and lysed in guanidinium isothiocyanate, while others were homogenized and lysed in phenol or in a suitable mixture of denaturants, such as TRIZOL (Life Technologies), a monophasic solution of phenol and guanidine isothiocyanate. The resulting lysates were centrifuged over CsCl cushions or extracted with chloroform. RNA was precipitated with either isopropanol or sodium acetate and ethanol, or by other routine methods.
  • RNA was provided with RNA and constructed the corresponding cDNA Ubraries. Otherwise, cDNA was synthesized and cDNA Ubraries were constructed with the UNLZAP vector system (Stratagene Cloning Systems, Inc. (Stratagene), La Jolla CA) or SUPERSCRIPT plasmid system (Life Technologies), using the recommended procedures or similar methods known in the art. (See, e.g., Ausubel, 1997, supra. Chapters 5.1 through 6.6.) Reverse transcription was initiated using oligo d(T) or random primers. Synthetic oligonucleotide adapters were Ugated to double stranded cDNA, and the cDNA was digested with the appropriate restriction enzyme or enzymes.
  • cDNA was size-selected (300-1000 bp) using SEPHACRYL S1000, SEPHAROSE CL2B, or SEPHAROSE CL4B column chromatography (Amersham Pharmacia Biotech) or preparative agarose gel electtophoresis.
  • cDN A were Ugated into compatible restriction enzyme sites of the polylinker of a suitable plasmid, e.g., PBLUESCRIPT plasmid (Stratagene), pSPORTl plasmid 5 (Life Technologies), or pINCY (Incyte).
  • Recombinant plasmids were transformed into competent E. coU cells including XLl-Blue, XLl-BlueMRF, or SOLR from Stratagene or DH5 ⁇ , DH10B, or ElectroMAX DH10B from Life Technologies.
  • Plasmids were purified using at least one of the following: the Magic or WIZARD Minipreps DNA purification system (Promega); the AGTC Miniprep purification kit (Edge BioSystems, Gaithersburg MD); and the QIAWELL 8, QIAWELL 8 Plus, and QIAWELL 8 Ultra plasmid purification systems or the R.E. A.L. PREP 96 plasmid purification kit (QIAGEN). Following 5 precipitation, plasmids were resuspended in 0.1 ml of distilled water and stored, with or without lyophiUzation, at 4°C
  • plasmid DNA was amplified from host cell lysates using direct Unk PCR in a high-throughput format. (Rao, V.B. (1994) Anal. Biochem. 216:1-14.) Host cell lysis and thermal cycling steps were carried out in a single reaction mixture. Samples were processed and stored in 384- o well plates, and the concentration of ampUfied plasmid DNA was quantified fluorometticaUy using
  • PICOGREEN dye (Molecular Probes, Inc. (Molecular Probes), Eugene OR) and a FLUOROSKAN II fluorescence scanner (Labsystems Oy, Helsinki, Finland).
  • cDNA sequencing reactions were processed using standard methods or high-throughput instrumentation such as the ABI CATALYST 800 thermal cycler (PE Biosystems) or the PTC-200 thermal cycler (MJ Research) in conjunction with the HYDRA microdispenser (Robbins Scientific Co ⁇ ., Sunnyvale CA) or the MICROLAB 2200 Uquid transfer system (Hamilton).
  • cDNA sequencing reactions were prepared using reagents provided by Amersham Pharmacia Biotech or suppUed in ABI o sequencing kits such as the ABI PRISM BIGDYE Terminator cycle sequencing ready reaction kit (PE
  • Component sequences from chromatograms were subject to PHRED analysis and assigned a quaUty score.
  • the sequences having at least a required quatity score were subject to various preprocessing editing pathways to eUminate, e.g., low quaUty 3' ends, vector and linker sequences, polyA tails, Au repeats, mitochondrial and ribosomal sequences, bacterial contamination sequences, and sequences smaller than 50 base pairs.
  • low-information sequences and repetitive elements e.g., dinucleotide repeats, Au repeats, etc.
  • sequences were then subject to assembly procedures in which the sequences were assigned to gene bins (bins). Each sequence could only belong to one bin. Sequences in each gene bin were assembled to produce consensus sequences (templates). Subsequent new sequences were added to existing bins using BLASTn (v.1.4 WashU) and CROSSMATCH. Candidate pairs were identified as all BLAST hits having a quaUty score greater than or equal to 150. Aignments of at least 82% local identity were accepted into the bin. The component sequences from each bin were assembled using a version of PHRAP. Bins with several overlapping component sequences were assembled using DEEP PHRAP.
  • each assembled template was determined based on the number and orientation of its component sequences. Template sequences as disclosed in the sequence Usting correspond to sense strand sequences (the "forward" reading frames), to the best determination. The complementary (antisense) strands are inherently disclosed herein.
  • the component sequences which were used to assemble each template consensus sequence are listed in Table 4, along with their positions along the template nucleotide sequences.
  • Bins were compared against each other and those having local similarity of at least 82% were combined and reassembled. Reassembled bins having templates of insufficient overlap (less than 95% local identity) were re-split. Asembled templates were also subject to analysis by STITCHER/EXON MAPPER algorithms which analyze the probabilities of the presence of spUce variants, alternatively spUced exons, spUce junctions, differential expression of alternative spliced genes across tissue types or disease states, etc. These resulting bins were subject to several rounds of the above assembly procedures.
  • bins were clone joined based upon clone informatioa If the 5' sequence of one clone was present in one bin and the 3' sequence from the same clone was present in a different bin, it was Ukely that the two bins actually belonged together in a single bin. The resulting combined bins underwent assembly procedures to regenerate the consensus sequences.
  • the template sequences were further analyzed by translating each template in all three forward reading frames and searching each translation against the Pfam database of hidden Markov model- based protein famiUes and domains using the HMMER software package (available to the pubUc from Washington University School of Medicine, St. Louis MO). Regions of templates which, when 5 translated, contain similarity to Pfam consensus sequences are reported in Table 2, along with descriptions of Pfam protein domains and families. Only those Pfam hits with an E-value of ⁇ 1 x IO "3 are reported. (See also World Wide Web site http://pfarawustl.edu/ for detailed descriptions of Pfam protein domains and famiUes.)
  • HMMER analysis as reported in Tables 2 and 3 may support the results of BLAST analysis as reported in Table 1 or may suggest alternative or additional properties of template- encoded polypeptides not previously uncovered by BLAST or other analyses.
  • Template sequences are further analyzed using the bioinformatics tools Usted in Table 6, or using sequence analysis software known in the art such as MACDNASIS PRO software (Hitachi Software Engineering, South San Francisco CA) and LASERGENE software (DNASTAR). Template sequences may be further queried against pubUc databases such as the GenBank rodent, mammaUan, vertebrate, prokaryote, and eukaryote databases.
  • Northern analysis is a laboratory technique used to detect the presence of a transcript of a gene and involves the hybridization of a labeled nucleotide sequence to a membrane on which RN from a particular cell type or tissue have been bound. (See, e.g., Sambrook, supra, ch. 7; Ausubel, 1995, supra, ch. 4 and 16.)
  • the product score takes into account both the degree of similarity between two sequences and the length of the sequence match.
  • the product score is a normaUzed value between 0 and 100, and is calculated as follows: the BLAST score is multipUed by the percent nucleotide identity and the product is divided by (5 times the length of the shorter of the two sequences).
  • the BLAST score is calculated by assigning a score of +5 for every base that matches in a high-scoring segment pair (HSP), and -4 for every mismatch. Two sequences may share more than one HSP (separated by gaps). If there is more than one HSP, then the pair with the highest BLAST score is used to calculate the product score.
  • the product score represents a balance between fractional overlap and quality in a BLAST aUgnment. For example, a product score of 100 is produced only for 100% identity over the entire length of the shorter of the two sequences being compared. A product score of 70 is produced either by 100% identity and 70% overlap at one end, or by 88% identity and 100% overlap at the other. A product score of 50 is produced either by 100% identity and 50% overlap at one end, or 79% identity and 100% overlap.
  • a tissue distribution profile is determined for each template by compiling the cDNA library tissue classifications of its component cDNA sequences.
  • Each component sequence is derived from a cDNA Ubrary constructed from a human tissue.
  • Each human tissue is classified into one of the following categories: cardiovascular system; connective tissue; digestive system; embryonic structures; endocrine system; exocrine glands; genitaUa, female; genitaUa, male; germ cells; hemic and immune system; liver; musculoskeletal system; nervous system; pancreas; respiratory system; sense organs; skin; stomatognathic system; unclassified/mixed; or urinary tract.
  • Template sequences, component sequences, and cDNA library/tissue information are found in the LIFESEQ GOLD database (Incyte Genomics, Palo Alto C A).
  • Table 5 shows the tissue distribution profile for the templates of the invention. For each template, the three most frequently observed tissue categories are shown in column 3, along with the percentage of component sequences belonging to each category. Only tissue categories with percentage values of > 10% are shown. A tissue distribution of "widely distributed" in column 3 indicates percentage values of ⁇ 10% in aU tissue categories.
  • Transcript images are generated as described in Seilhamer et al., "Comparative Gene Transcript Analysis," U.S. Patent Number 5,840,484, inco ⁇ orated herein by reference.
  • Oligonucleotide primers designed using a dithp of the Sequence Listing are used to extend the nucleic acid sequence.
  • One primer is synthesized to initiate 5' extension of the template, and the other primer, to initiate 3' extension of the template.
  • the initial primers may be designed using OLIGO 4.06 software (National Biosciences, Inc. (National Biosciences), Plymouth MN), or another appropriate program, to be about 22 to 30 nucleotides in length, to have a GC content of about 50% or more, and to anneal to the target sequence at temperatures of about 68 °C to about 72 °C Any stretch of nucleotides which would result in hai ⁇ in structures and primer-primer dimerizations are avoided.
  • Selected human cDNA libraries are used to extend the sequence. If more than one extension is necessary or desired, additional or nested sets of primers are designed.
  • PCR is performed in 96-well plates using the PTC-200 thermal cycler (MJ Research).
  • the reaction mix 5 contains DNA template, 200 nmol of each primer, reaction buffer containing Mg 2+ , (NH 4 ) 2 S0 4 , and ⁇ - mercaptoethanol, Taq DNA polymerase (Amersham Pharmacia Biotech), ELONGASE enzyme (Life Technologies), and Pfu DNA polymerase (Stratagene), with the following parameters for primer pair PCI A and PCI B: Step 1 : 94°C, 3 min; Step 2: 94°C, 15 sec; Step 3: 60°C, 1 min; Step 4: 68°C, 2 min; Step 5: Steps 2, 3, and 4 repeated 20 times; Step 6: 68°C, 5 min; Step 7: storage at 4°C In the 0 alternative, the parameters for primer pair T7 and SK+ are as follows: Step 1 : 94°C, 3 min;
  • the concentration of DNA in each well is determined by dispensing 100 ⁇ l PICOGREEN quantitation reagent (0.25% (v/v); Molecular Probes) dissolved in IX Tris-EDTA (TE) and 0.5 ⁇ l of 5 undiluted PCR product into each well of an opaque fluorimeter plate (Corning Inco ⁇ orated (Corning), Corning NY), allowing the DNA to bind to the reagent.
  • the plate is scanned in a FLUOROSKAN II (Labsystems Oy) to measure the fluorescence of the sample and to quantify the concentration of DNA
  • FLUOROSKAN II (Labsystems Oy) to measure the fluorescence of the sample and to quantify the concentration of DNA
  • a 5 ⁇ l to 10 ⁇ l ahquot of the reaction mixture is analyzed by electrophoresis on a 1 % agarose mini-gel to determine which reactions are successful in extending the sequence.
  • the extended nucleotides are desalted and concentrated, transferred to 384-well plates, digested with CviJI cholera virus endonuclease (Molecular Biology Research, Madison WI), and sonicated or sheared prior to reUgation into pUC 18 vector (Amersham Pharmacia Biotech).
  • the digested nucleotides are separated on low concentration (0.6 to 0.8%) agarose gels, fragments are excised, and agar digested with AGAR ACE (Promega). Extended clones are 5 religated using T4 Ugase (New England Biolabs, Inc., Beverly MA) into pUC 18 vector (Amersham Pharmacia Biotech), treated with Pfu DNA polymerase (Sttatagene) to fill-in restriction site overhangs, and transfected into competent E. coh cells. Transformed cells are selected on antibiotic-containing media, individual colonies are picked and cultured overnight at 37 °C in 384-well plates in LB/2x carbenicilUn Uquid media. o The cells are lysed, and DNA is ampUfied by PCR using Taq DNA polymerase (Amersham
  • Step 1 94°C, 3 min
  • Step 2 94°C, 15 sec
  • Step 3 60°C, 1 min
  • Step 4 72°C, 2 min
  • Step 5 steps 2, 3, and 4 repeated 29 times
  • Step 6 72°C, 5 min
  • Step 7 storage at 4°C DNA is quantified by PICOGREEN reagent (Molecular Probes) as described above. Samples with low DNA recoveries are reampUfied using the same conditions as described above.
  • Samples are diluted with 20% dimethysulfoxide (1 :2, v/v), and sequenced using DYENAMIC energy transfer sequencing primers and the DYENAMIC DIRECT kit (Amersham Pharmacia Biotech) or the ABI PRISM BIGDYE Terminator cycle sequencing ready reaction kit (PE Biosystems). 5
  • the dithp is used to obtain regulatory sequences (promoters, introns, and enhancers) using the procedure above, oUgonucleotides designed for such extension, and an appropriate genomic Ubrary.
  • Hybridization probes derived from the dithp of the Sequence Listing are employed for screening cDN , mRNA, or genomic DNA.
  • the labehng of probe nucleotides between 100 and 1000 nucleotides in length is specifically described, but essentially the same procedure may be used with larger cDNA fragments.
  • Probe sequences are labeled at room temperature for 30 minutes using a T4 polynucleotide kinase, ⁇ P-ATP, and 0.5X One-Phor-Al Plus (Amersham Pharmacia Biotech) 5 buffer and purified using a ProbeQuant G-50 Microcolumn (Amersham Pharmacia Biotech). The probe mixture is diluted to IO 7 dpm/ ⁇ g/ml hybridization buffer and used in a typical membrane-based hybridization analysis.
  • the DNA is digested with a restriction endonuclease such as Eco RV and is electrophoresed through a 0.7% agarose gel.
  • the DNA fragments are transferred from the agarose to nylon membrane o (NYTRAN Plus, Schleicher & Schuell, Inc., Keene NH) using procedures specified by the manufacturer of the membrane.
  • Prehybridization is carried out for three or more hours at 68 °C, and hybridization is carried out overnight at 68 °C
  • blots are sequentially washed at room temperature under increasingly stringent conditions, up to 0. lx saUne sodium citrate (SSC) and 0.5% sodium dodecyl sulfate. After the blots are placed in a PHOSPHORIMAGER cassette 5 (Molecular Dynamics) or are exposed to autoradiography film, hybridization patterns of standard and experimental lanes are compared. Essentially the same procedure is employed when screening RNA
  • the cDNA sequences which were used to assemble SEQ ID NO: 1-71 are compared with o sequences from the Incyte LIFESEQ database and pubUc domain databases using BLAST and other implementations of the Smith- Waterman algorithm. Sequences from these databases that match SEQ ID NO: 1-71 are assembled into clusters of contiguous and overlapping sequences using assembly algorithms such as PHRAP (Table 6). Radiation hybrid and genetic mapping data available from public resources such as the Stanford Human Genome Center (SHGC), Whitehead Institute for Genome Research (WIGR), and G ⁇ nethon are used to determine if any of the clustered sequences have been previously mapped.
  • SHGC Stanford Human Genome Center
  • WIGR Whitehead Institute for Genome Research
  • G ⁇ nethon are used to determine if any of the clustered sequences have been previously mapped.
  • a mapped sequence in a cluster will result in the assignment of all sequences of that cluster, including its particular SEQ ID NO:, to that map location.
  • the genetic map locations of SEQ ID NO: 1-71 are described as ranges, or intervals, of human chromosomes.
  • the map 5 position of an interval, in centiMorgans, is measured relative to the terminus of the chromosome's p- arm. (The centiMorgan (cM) is a unit of measurement based on recombination frequencies between chromosomal markers.
  • cM is roughly equivalent to 1 megabase (Mb) of DNA in humans, although this can vary widely due to hot and cold spots of recombination.
  • Mb megabase
  • the cM distances are based on genetic markers mapped by Gen ⁇ thon which provide boundaries for radiation hybrid o markers whose sequences were included in each of the clusters.
  • RNA is isolated from tissue samples using the guanidinium thiocyanate method and 5 polyA + RNA is purified using the oligo (dT) cellulose method.
  • Each polyA + RNA sample is reverse transcribed using MMLV reverse-transcriptase, 0.05 pg/ ⁇ l oUgo-dT primer (21mer), IX first sfrand buffer, 0.03 units/ ⁇ l RNase inhibitor, 500 ⁇ M dATP, 500 ⁇ M dGTP, 500 ⁇ M dTTP, 40 ⁇ M dCTP, 40 ⁇ M dCTP-Cy3 (BDS) or dCTP-Cy5 (Amersham Pharmacia Biotech).
  • the reverse transcription reaction is performed in a 25 ml volume containing 200 ng polyA + RNA with GEMB RIGHT kits o (Incyte).
  • Specific control polyA + RNA are synthesized by in vitro transcription from non-coding yeast genomic DNA (W. Lei, unpubhshed).
  • a quantitative controls, the control mRNA at 0.002 ng, 0.02 ng, 0.2 ng, and 2 ng are diluted into reverse ttanscription reaction at ratios of 1 :100,000, 1 :10,000, 1 : 1000, 1 :100 (w/w) to sample mRNA respectively.
  • control mRNA are diluted into reverse transcription reaction at ratios of 1:3, 3:1, 1:10, 10:1, 1:25, 25:1 (w/w) to sample mRNA differential 5 expression patterns. After incubation at 37° C for 2 hr, each reaction sample (one with Cy3 and another with Cy5 labeUng) is treated with 2.5 ml of 0.5M sodium hydroxide and incubated for 20 minutes at 85° C to the stop the reaction and degrade the RNA. Probes are purified using two successive CHROMA SPIN 30 gel filtration spin columns (CLONTECH Laboratories, Inc.
  • reaction samples are ethanol precipitated using 1 ml of glycogen (1 o mg/ml), 60 ml sodium acetate, and 300 ml of 100% ethanol.
  • the probe is then dried to completion using a SpeedVAC (Savant Instruments Inc., Holbrook NY) and resuspended in 14 ⁇ l 5X SSC/0.2% SDS.
  • Microarray Preparation Sequences of the present invention are used to generate array elements.
  • Each array element is ampUfied from bacterial cells containing vectors with cloned cDNA inserts.
  • PCR ampUfication uses primers complementary to the vector sequences flanking the cDNA insert.
  • Array elements are ampUfied in thirty cycles of PCR from an initial quantity of 1-2 ng to a final quantity greater than 5 ⁇ g.
  • AmpUfied array elements are then purified using SEPHACRYL-400 (Amersham Pharmacia Biotech). Purified array elements are immobiUzed on polymer-coated glass sUdes.
  • Glass microscope sUdes are cleaned by ultrasound in 0.1% SDS and acetone, with extensive distilled water washes between and after treatments. Glass slides are etched in 4% hydrofluoric acid (VWR Scientific Products Corporation (VWR), West Chester, PA), washed extensively in distilled water, and coated with 0.05% aminopropyl silane (Sigma) in 95% ethanol. Coated sUdes are cured in a 110°C oven. Array elements are appUed to the coated glass substrate using a procedure described in US Patent No. 5,807,522, inco ⁇ orated herein by reference.
  • Microarrays are UV-crosslinked using a STRATALINKER UV-crossUnker (Stratagene).
  • Microarrays are washed at room temperature once in 0.2% SDS and three times in distilled water. Non-specific binding sites are blocked by incubation of microarrays in 0.2% casein in phosphate buffered saUne (PBS) (Tropix, Inc., Bedford, MA) for 30 minutes at 60° C followed by washes in 0.2% SDS and distilled water as before.
  • PBS phosphate buffered saUne
  • Hybridization reactions contain 9 ⁇ l of probe mixture consisting of 0.2 ⁇ g each of Cy3 and Cy5 labeled cDNA synthesis products in 5X SSC, 0.2% SDS hybridization buffer.
  • the probe mixture is heated to 65° C for 5 minutes and is aliquoted onto the microarray surface and covered with an 1.8 cm 2 coversUp.
  • the arrays are ttansferred to a wate ⁇ roof chamber having a cavity just sUghtly larger than a microscope sUde.
  • the chamber is kept at 100% humidity internally by the addition of 140 ⁇ l of 5x SSC in a corner of the chamber.
  • the chamber containing the arrays is incubated for about 6.5 hours at 60° C.
  • the arrays are washed for 10 min at 45° C in a first wash buffer (IX SSC, 0.1% SDS), three times for 10 minutes each at 45° C in a second wash buffer (0. IX SSC), and dried.
  • Reporter-labeled hybridization complexes are detected with a microscope equipped with an Lnnova 70 mixed gas 10 W laser (Coherent, Inc., Santa Clara CA) capable of generating spectral lines at 488 nm for excitation of Cy3 and at 632 nm for excitation of Cy5.
  • the excitation laser Ught is focused on the array using a 20X microscope objective (Nikon, Inc., Melville NY).
  • the sUde containing the array is placed on a computer-controlled X-Y stage on the microscope and raster- scanned past the objective.
  • the 1.8 cm x 1.8 cm array used in the present example is scanned with a resolution of 20 micrometers. 5 In two separate scans, a mixed gas multiline laser excites the two fluorophores sequentially.
  • Emitted light is spUt, based on wavelength, into two photomultipUer tube detectors (PMT R1477, Hamamatsu Photonics Systems, Bridgewater NJ) corresponding to the two fluorophores.
  • Appropriate filters positioned between the array and the photomultipUer tubes are used to filter the signals.
  • the emission maxima of the fluorophores used are 565 nm for Cy3 and 650 nm for Cy5.
  • Each array is o typically scanned twice, one scan per fluorophore using the appropriate filters at the laser source, although the apparatus is capable of recording the spectra from both fluorophores simultaneously.
  • the sensitivity of the scans is typically caUbrated using the signal intensity generated by a cDNA control species added to the probe mix at a known concentration.
  • a specific location on the array contains a complementary DNA sequence, allowing the intensity of the signal at that location to 5 be correlated with a weight ratio of hybridizing species of 1 : 100,000.
  • the calibration is done by labehng samples of the calibrating cDNA with the two fluorophores and adding identical amounts of each to the hybridization mixture. o
  • the output of the photomultipUer tube is digitized using a 12-bit RTI-835H analog-to-digital
  • A/D conversion board Analog Devices, Inc., Norwood, MA
  • the digitized data are displayed as an image where the signal intensity is mapped using a Unear 20-color transformation to a pseudocolor scale ranging from blue (low signal) to red (high signal).
  • the data is also analyzed quantitatively. Where two different fluorophores are excited and 5 measured simultaneously, the data are first corrected for optical crosstalk (due to overlapping emission spectra) between the fluorophores using each fluorophore's emission spectrum.
  • a grid is superimposed over the fluorescence signal image such that the signal from each spot is centered in each element of the grid.
  • the fluorescence signal within each element is then integrated to obtain a numerical value corresponding to the average intensity of the signal.
  • the software used for o signal analysis is the GEMTOOLS gene expression analysis program (Incyte).
  • Sequences complementary to the dithp are used to detect, decrease, or inhibit expression of the naturally occurring nucleotide.
  • the use of oUgonucleotides comprising from about 15 to 30 base pairs is typical in the art. However, smaller or larger sequence fragments can also be used.
  • Appropriate oUgonucleotides are designed from the dithp using OLIGO 4.06 software (National Biosciences) or other appropriate programs and are synthesized using methods standard in the art or ordered from a commercial suppUer.
  • a complementary oUgonucleotide is designed from the 5 most unique 5 ' sequence and used to prevent transcription factor binding to the promoter sequence.
  • To inhibit translation, a complementary oligonucleotide is designed to prevent ribosomal binding and processing of the transcript.
  • DITHP expression and purification of DITHP is accomplished using bacterial or virus-based expression systems.
  • cDNA is subcloned into an appropriate vector containing an antibiotic resistance gene and an inducible promoter that directs high levels of cDNA ttanscription.
  • promoters include, but are not Umited to, the trp-lac (tac) hybrid promoter and the T5 or T7 bacteriophage promoter in conjunction with the lac operator 5 regulatory element.
  • Recombinant vectors are transformed into suitable bacterial hosts, e.g.,
  • DITHP upon induction with isopropyl beta-D- thiogalactopyranoside (IPTG).
  • IPTG isopropyl beta-D- thiogalactopyranoside
  • Expression of DITHP in eukaryotic cells is achieved by infecting insect or mammaUan cell Unes with recombinant Autographica caUfornica nuclear polyhedrosis virus (AcMNPV), commonly known as baculovirus.
  • AcMNPV Autographica caUfornica nuclear polyhedrosis virus
  • the nonessential polyhedrin gene of baculovirus is o replaced with cDNA encoding DITHP by either homologous recombination or bacterial-mediated transposition involving transfer plasmid intermediates. Viral infectivity is maintained and the strong polyhedrin promoter drives high levels of cDNA ttanscription.
  • baculovirus Recombinant baculovirus is used to infect Spodoptera frugiperda (Sf9) insect cells in most cases, or human hepatocytes, in some cases. Infection of the latter requires additional genetic modifications to baculovirus. (See e.g., Engelhard, 5 supra; and Sandig, supra.)
  • DITHP is synthesized as a fusion protein with, e.g., glutathione S- fransferase (GST) or a peptide epitope tag, such as FLAG or 6-His, permitting rapid, single-step, affinity-based purification of recombinant fusion protein from crude cell lysates.
  • GST a 26-kilodalton enzyme from Schistosoma iaponicum. enables the purification of fusion proteins on immobilized o glutathione under conditions that maintain protein activity and antigenicity (Amersham Pharmacia).
  • the GST moiety can be proteolytically cleaved from DITHP at specifically engineered sites.
  • FLAG an 8-amino acid peptide
  • 6-His a stretch of six consecutive histidine residues, enables purification on metal-chelate resins (QIAGEN). Methods for protein expression and purification are discussed in Ausubel (1995, supra. Chapters 10 and 16). Purified DITHP obtained by these methods can be used directly in the following activity assay.
  • DITHP activity is demonstrated through a variety of specific assays, some of which are outUned below.
  • Oxidoreductase activity of DITHP is measured by the increase in extinction coefficient of NAD(P)H coenzyme at 340 nmfor the measurement of oxidation activity, or the decrease in extinction l o coefficient of NAD(P)H coenzyme at 340 nmfor the measurement of reduction activity (Dalziel, K.

Landscapes

  • Enzymes And Modification Thereof (AREA)
  • Medicines That Contain Protein Lipid Enzymes And Other Medicines (AREA)

Abstract

The present invention provides purified human polynucleotides for diagnostics and therapeutics (dithp). Also encompassed are the polypeptides (DITHP) encoded by dithp. The invention also provides for the use of dithp, or complements, oligonucleotides, or fragments thereof in diagnostic assays. The invention further provides for vectors and host cells containing dithp for the expression of DITHP. The invention additionally provides for the use of isolated and purified DITHP to induce antibodies and to screen libraries of compounds and the use of anti-DITHP antibodies in diagnostic assays. Also provided are microarrays containing dithp and methods of use.

Description

MOLECULES FOR DIAGNOSTICS AND THERAPEUTICS
TECHNICAL FIELD
The present invention relates to human molecules and to the use of these sequences in the 5 diagnosis, study, prevention, and treatment of diseases associated with, as well as effects of exogenous compounds on, the expression of human molecules.
BACKGROUND OF THE INVENTION
The human genome is comprised of thousands of genes, many encoding gene products that o function in the maintenance and growth of the various cells and tissues in the body. Aberrant expression or mutations in these genes and their products is the cause of, or is associated with, a variety of human diseases such as cancer and other cell proliferative disorders, autoimmune/inflammatory disorders, infections, developmental disorders, endocrine disorders, metabolic disorders, neurological disorders, gastrointestinal disorders, transport disorders, and connective tissue disorders. The 5 identification of these genes and their products is the basis of an ever-expanding effort to find markers for early detection of diseases, and targets for their prevention and treatment. Therefore, these genes and their products are useful as diagnostics and therapeutics. These genes may encode, for example, enzyme molecules, molecules associated with growth and development, biochemical pathway molecules, extracellular information transmission molecules, receptor molecules, intracellular signaling molecules, o membrane transport molecules, protein modification and maintenance molecules, nucleic acid synthesis and modification molecules, adhesion molecules, antigen recognition molecules, secreted and extracellular matrix molecules, cytoskeletal molecules, ribosomal molecules, electron transfer associated molecules, transcription factor molecules, chromatin molecules, cell membrane molecules, and organelle associated molecules. 5 For example, cancer represents a type of cell proliferative disorder that affects nearly every tissue in the body. A wide variety of molecules, either aberrantly expressed or mutated, can be the cause of, or involved with, various cancers because tissue growth involves complex and ordered patterns of cell proliferation, cell differentiation, and apoptosis. Cell proliferation must be regulated to maintain both the number of cells and their spatial organization. This regulation depends upon the o appropriate expression of proteins which control cell cycle progression in response to extracellular signals such as growth factors and other mitogens, and intracellular cues such as DNA damage or nutrient starvation. Molecules which directly or indirectly modulate cell cycle progression fall into several categories, including growth factors and their receptors, second messenger and signal transduction proteins, oncogene products, tumor-suppressor proteins, and mitosis-promoting factors. Aberrant expression or mutations in any of these gene products can result in cell proliferative disorders such as cancer. Oncogenes are genes generally derived from normal genes that, through abnormal expression or mutation, can effect the transformation of a normal cell to a malignant one (oncogenesis). Oncoproteins, encoded by oncogenes, can affect cell proliferation in a variety of ways and include growth factors, growth factor receptors, intracellular signal transducers, nuclear transcription factors, and cell-cycle control proteins. In contrast, tumor-suppressor genes are involved in inhibiting cell proliferation. Mutations which cause reduced function or loss of function in tumor-suppressor genes result in aberrant cell proliferation and cancer. Although many different genes and their products have been found to be associated with cell proliferative disorders such as cancer, many more may exist that are yet to be discovered.
DNA-based arrays can provide a simple way to explore the expression of a single polymorphic gene or a large number of genes. When the expression of a single gene is explored, DNA-based arrays are employed to detect the expression of specific gene variants. For example, a p53 tumor suppressor gene array is used to determine whether individuals are carrying mutations that predispose them to cancer. A cytochrome p450 gene array is useful to determine whether individuals have one of a number of specific mutations that could result in increased drug metabolism, drug resistance or drug toxicity. DNA-based array technology is especially relevant for the rapid screening of expression of a large number of genes. There is a growing awareness that gene expression is affected in a global fashion. A genetic predisposition, disease or therapeutic treatment may affect, directly or indirectly, the expression of a large number of genes. In some cases the interactions may be expected, such as when the genes are part of the same signaling pathway. In other cases, such as when the genes participate in separate signaling pathways, the interactions may be totally unexpected. Therefore, DNA-based arrays can be used to investigate how genetic predisposition, disease, or therapeutic treatment affects the expression of a large number of genes.
Enzyme Molecules
SEQ ID NO:l, SEQ ID NO:2, SEQ ID NO:3, SEQ ID NO:4, SEQ ID NO:5, SEQ ID NO:6, SEQ ID NO:7, and SEQ ID NO:8 encode, for example, human enzyme molecules.
The cellular processes of biogenesis and biodegradation involve a number of key enzyme classes including oxidoreductases, transferases, hydrolases, lyases, isomerases, and ligases. These enzyme classes are each comprised of numerous substrate-specific enzymes having precise and well regulated functions. These enzymes function by facilitating metabolic processes such as glycolysis, the tricarboxylic cycle, and fatty acid metabolism; synthesis or degradation of amino acids, steroids, phospholipids, alcohols, etc.; regulation of cell signalling, proliferation, inflamation, apoptosis, etc., and through catalyzing critical steps in DNA replication and repair, and the process of translation. Oxidoreductases
Many pathways of biogenesis and biodegradation require oxidoreductase (dehydrogenase or reductase) activity, coupled to the reduction or oxidation of a donor or acceptor cofactor. Potential cofactors include cytochromes, oxygen, disulfide, iron-sulfur proteins, flavin adenine dinucleotide (FAD), and the nicotinamide adenine dinucleotides NAD and NADP (Newsholme, E.A. and A.R. Leech (1983) Biochemistry for the Medical Sciences, John Wiley and Sons, Chichester, U.K., pp. 779-793). Reductase activity catalyzes the transfer of electrons between substrate(s) and cofactor(s) with concurrent oxidation of the cofactor. The reverse dehydrogenase reaction catalyzes the reduction of a cofactor and consequent oxidation of the substrate. Oxidoreductase enzymes are a broad superfamily of proteins that catalyze numerous reactions in all cells of organisms ranging from bacteria to plants to humans. These reactions include metabolism of sugar, certain detoxification reactions in the liver, and the synthesis or degradation of fatty acids, amino acids, glucocorticoids, estrogens, androgens, and prostaglandins. Different family members are named according to the direction in which their reactions are typically catalyzed; thus they may be referred to as oxidoreductases, oxidases, reductases, or dehydrogenases. In addition, family members often have distinct cellular localizations, including the cytosol, the plasma membrane, mitochondrial inner or outer membrane, and peroxisomes.
Short-chain alcohol dehydrogenases (SCADs) are a family of dehydrogenases that only share 15% to 30% sequence identity, with similarity predominantly in the coenzyme binding domain and the substrate binding domain. In addition to the well-known role in detoxification of ethanol, SCADs are also involved in synthesis and degradation of fatty acids, steroids, and some prostaglandins, and are therefore implicated in a variety of disorders such as lipid storage disease, myopathy, SCAD deficiency, and certain genetic disorders. For example, retinol dehydrogenase is a SCAD-family member (Simon, A. et al. (1995) J. Biol. Chem. 270:1107-1112) that converts retinol to retinal, the precursor of retinoic acid. Retinoic acid, a regulator of differentiation and apoptosis, has been shown to down-regulate genes involved in cell proliferation and inflammation (Chai, X. et al. (1995) J. Biol. Chem. 270:3900-3904). In addition, retinol dehydrogenase has been linked to hereditary eye diseases such as autosomal recessive childhood-onset severe retinal dystrophy (Simon, A. et al. (1996) Genomics 36:424-430).
Propagation of nerve impulses, modulation of cell proliferation and differentiation, induction of the immune response, and tissue homeostasis involve neurotransmitter metabolism (Weiss, B. (1991) Neurotoxicology 12:379-386; Collins, S.M. et al. (1992) Ann. N.Y. Acad. Sci. 664:415-424; Brown, J.K. and H. Imam (1991) J. Inherit. Metab. Dis. 14:436-458). Many pathways of neurotransmitter metabolism require oxidoreductase activity, coupled to reduction or oxidation of a cofactor, such as NAD7NADH (Newsholme, E.A. and A.R. Leech (1983) Biochemistry for the Medical Sciences, John Wiley and Sons, Chichester, U.K. pp. 779-793). Degradation of catecholamines (epinephrine or norepinephrine) requires alcohol dehydrogenase (in the brain) or aldehyde dehydrogenase (in peripheral tissue). NAD+ -dependent aldehyde dehydrogenase oxidizes 5- 5 hydroxyindole-3-acetate (the product of 5-hydroxytryptamine (serotonin) metabolism) in the brain, blood platelets, liver and pulmonary endothelium (Newsholme, supra, p. 786). Other neurotransmitter degradation pathways that utilize NADVNADH-dependent oxidoreductase activity include those of L-DOPA (precursor of dopamine, a neuronal excitatory compound), glycine (an inhibitory neurotransmitter in the brain and spinal cord), histamine (liberated from mast cells during o the inflammatory response), and taurine (an inhibitory neurotransmitter of the brain stem, spinal cord and retina) (Newsholme, supra, pp. 790, 792). Epigenetic or genetic defects in neurotransmitter metabolic pathways can result in a spectrum of disease states in different tissues including Parkinson disease and inherited myoclonus (McCance, K.L. and S.E. Huether (1994) Pathophysiology. Mosby- Year Book, Inc., St. Louis MO, pp. 402-404; Gundlach, A.L. (1990) FASEB J. 4:2761-2766). 5 Tetrahydrofolate is a derivatized glutamate molecule that acts as a carrier, providing activated one-carbon units to a wide variety of biosynthetic reactions, including synthesis of purines, pyrimidines, and the amino acid methionine. Tetrahydrofolate is generated by the activity of a holoenzyme complex called tetrahydrofolate synthase, which includes three enzyme activities: tetrahydrofolate dehydrogenase, tetrahydrofolate cyclohydrolase, and tetrahydrofolate synthetase. o Thus, tetrahydrofolate dehydrogenase plays an important role in generating building blocks for nucleic and amino acids, crucial to proliferating cells.
3-Hydroxyacyl-CoA dehydrogenase (3HACD) is involved in fatty acid metabolism. It catalyzes the reduction of 3-hydroxyacyl-CoA to 3-oxoacyl-CoA, with concomitant oxidation of NAD to NADH, in the mitochondria and peroxisomes of eukaryotic cells. In peroxisomes, 3HACD 5 and enoyl-CoA hydratase form an enzyme complex called bifunctional enzyme, defects in which are associated with peroxisomal bifunctional enzyme deficiency. This interruption in fatty acid metabolism produces accumulation of very-long chain fatty acids, disrupting development of the brain, bone, and adrenal glands. Infants born with this deficiency typically die within 6 months (Watkins, P. et al. (1989) J. Clin. Invest. 83:771-777; Online Mendelian Inheritance in Man (OMIM), 0 #261515). The neurodegeneration that is characteristic of Alzheimer's disease involves development of extracellular plaques in certain brain regions. A major protein component of these plaques is the peptide amyloid-β (Aβ), which is one of several cleavage products of amyloid precursor protein (APP). 3HACD has been shown to bind the Aβ peptide, and is overexpressed in neurons affected in Alzheimer's disease. In addition, an antibody against 3HACD can block the toxic effects of Aβ in a 5 cell culture model of Alzheimer's disease (Yan, S. et al. (1997) Nature 389:689-695; OMIM, #602057).
Steroids, such as estrogen, testosterone, corticosterone, and others, are generated from a common precursor, cholesterol, and are interconverted into one another. A wide variety of enzymes act upon cholesterol, including a number of dehydrogenases. Steroid dehydrogenases, such as the 5 hydroxysteroid dehydrogenases, are involved in hypertension, fertility, and cancer (Duax, W.L. and D. Ghosh (1997) Steroids 62:95-100). One such dehydrogenase is 3-oxo-5-α-steroid dehydrogenase (OASD), a microsomal membrane protein highly expressed in prostate and other androgen-responsive tissues. OASD catalyzes the conversion of testosterone into dihydrotestosterone, which is the most potent androgen. Dihydrotestosterone is essential for the formation of the male phenotype during o embryogenesis, as well as for proper androgen-mediated growth of tissues such as the prostate and male genitalia. A defect in OASD that prevents the conversion of testosterone into dihydrotestosterone leads to a rare form of male pseudohermaphroditis, characterized by defective formation of the external genitalia (Andersson, S. et al. (1991) Nature 354:159-161; Labrie, F. et al. (1992) Endocrinology 131:1571-1573; OMIM #264600). Thus, OASD plays a central role in sexual 5 differentiation and androgen physiology.
17β-hydroxysteroid dehydrogenase (17βHSD6) plays an important role in the regulation of the male reproductive hormone, dihydrotestosterone (DHTT). 17βHSD6 acts to reduce levels of DHTT by oxidizing a precursor of DHTT, 3α-diol, to androsterone which is readily glucuronidated and removed from tissues. 17βHSD6 is active with both androgen and estrogen substrates when 0 expressed in embryonic kidney 293 cells. At least five other isozymes of 17βHSD have been identified that catalyze oxidation and/or reduction reactions in various tissues with preferences for different steroid substrates (Biswas, M.G. and D.W. Russell (1997) J. Biol. Chem. 272:15959- 15966). For example, 17βHSDl preferentially reduces estradiol and is abundant in the ovary and placenta. 17βHSD2 catalyzes oxidation of androgens and is present in the endometrium and placenta. 5 17βHSD3 is exclusively a reductive enzyme in the testis (Geissler, W.M. et al. (1994) Nat. Genet.
7:34-39). An excess of androgens such as DHTT can contribute to certain disease states such as benign prostatic hyperplasia and prostate cancer.
Oxidoreductases are components of the fatty acid metabolism pathways in mitochondria and peroxisomes. The main beta-oxidation pathway degrades both saturated and unsaturated fatty acids, o while the auxiliary pathway performs additional steps required for the degradation of unsaturated fatty acids. The auxiliary beta-oxidation enzyme 2,4-dienoyl-CoA reductase catalyzes the removal of even-numbered double bonds from unsaturated fatty acids prior to their entry into the main beta- oxidation pathway. The enzyme may also remove odd-numbered double bonds from unsaturated fatty acids (Koivuranta, K.T. et al. (1994) Biochem. J. 304:787-792; Smeland, T.E. et al. (1992) Proc. 5 Natl. Acad. Sci. USA 89:6673-6677). 2,4-dienoyl-CoA reductase is located in both mitochondria and peroxisomes. Inherited deficiencies in mitochondrial and peroxisomal beta-oxidation enzymes are associated with severe diseases, some of which manifest themselves soon after birth and lead to death within a few years. Defects in beta-oxidation are associated with Reye's syndrome, Zellweger syndrome, neonatal adrenoleukodystrophy, infantile Refsum's disease, acyl-CoA oxidase deficiency, 5 and bifunctional protein deficiency (Suzuki, Y. et al. (1994) Am. J. Hum. Genet. 54:36-43; Hoefler, supra: Cotran, R.S. et al. (1994) Robbins Pathologic Basis of Disease. W.B. Saunders Co., Philadelphia PA, p.866). Peroxisomal beta-oxidation is impaired in cancerous tissue. Although neoplastic human breast epithelial cells have the same number of peroxisomes as do normal cells, fatty acyl-CoA oxidase activity is lower than in control tissue (el Bouhtoury, F. et al. (1992) J. Pathol. 0 166:27-35). Human colon carcinomas have fewer peroxisomes than normal colon tissue and have lower fatty-acyl-CoA oxidase and bifunctional enzyme (including enoyl-CoA hydratase) activities than normal tissue (Cable, S. et al. (1992) Virchows Arch. B Cell Pathol. Incl. Mol. Pathol. 62:221- 226). Another important oxidoreductase is isocitrate dehydrogenase, which catalyzes the conversion of isocitrate to a-ketoglutarate, a substrate of the citric acid cycle. Isocitrate dehydrogenase can be 5 either NAD or NADP dependent, and is found in the cytosol, mitochondria, and peroxisomes. Activity of isocitrate dehydrogenase is regulated developmentally, and by hormones, neurotransmitters, and growth factors.
Hydroxypyruvate reductase (HPR), a peroxisomal 2-hydroxyacid dehydrogenase in the glycolate pathway, catalyzes the conversion of hydroxypyruvate to glycerate with the oxidation of o both NADH and NADPH. The reverse dehydrogenase reaction reduces NAD+ and NADP+. HPR recycles nucleotides and bases back into pathways leading to the synthesis of ATP and GTP. ATP and GTP are used to produce DNA and RNA and to control various aspects of signal transduction and energy metabolism. Inhibitors of purine nucleotide biosynthesis have long been employed as antiproliferative agents to treat cancer and viral diseases. HPR also regulates biochemical synthesis 5 of serine and cellular serine levels available for protein synthesis.
The mitochondrial electron transport (or respiratory) chain is a series of oxidoreductase-type enzyme complexes in the mitochondrial membrane that is responsible for the transport of electrons from NADH through a series of redox centers within these complexes to oxygen, and the coupling of this oxidation to the synthesis of ATP (oxidative phosphorylation). ATP then provides the primary o source of energy for driving a cell's many energy-requiring reactions. The key complexes in the respiratory chain are NADH:ubiquinone oxidoreductase (complex I), succinate:ubiquinone oxidoreductase (complex II), cytochrome crb oxidoreductase (complex III), cytochrome c oxidase (complex IV), and ATP synthase (complex V) (Alberts, B. et al. (1994) Molecular Biology of the Cell, Garland Publishing, Inc., New York NY, pp. 677-678). All of these complexes are located on 5 the inner matrix side of the mitochondrial membrane except complex II, which is on the cytosolic side. Complex II transports electrons generated in the citric acid cycle to the respiratory chain. The electrons generated by oxidation of succinate to fumarate in the citric acid cycle are transferred through electron carriers in complex II to membrane bound ubiquinone (Q). Transcriptional regulation of these nuclear-encoded genes appears to be the predominant means for controlling the 5 biogenesis of respiratory enzymes. Defects and altered expression of enzymes in the respiratory chain are associated with a variety of disease conditions.
Other dehydrogenase activities using NAD as a cofactor are also important in mitochondrial function. 3-hydroxyisobutyrate dehydrogenase (3HBD), important in valine catabolism, catalyzes the NAD-dependent oxidation of 3-hydroxyisobutyrate to methylmalonate semialdehyde within o mitochondria. Elevated levels of 3-hydroxyisobutyrate have been reported in a number of disease states, including ketoacidosis, methylmalonic acidemia, and other disorders associated with deficiencies in methylmalonate semialdehyde dehydrogenase (Rougraff, P.M. et al. (1989) J. Biol. Chem. 264:5899-5903).
Another mitochondrial dehydrogenase important in amino acid metabolism is the enzyme 5 isovaleryl-CoA-dehydrogenase (IVD). IVD is involved in leucine metabolism and catalyzes the oxidation of isovaleryl-CoA to 3-methylcrotonyl-CoA. Human IVD is a tetrameric flavoprotein that is encoded in the nucleus and synthesized in the cytosol as a 45 kDa precursor with a mitochondrial import signal sequence. A genetic deficiency, caused by a mutation in the gene encoding IVD, results in the condition known as isovaleric acidemia. This mutation results in inefficient mitochondrial 0 import and processing of the IVD precursor (Vockley, J. et al. (1992) J. Biol. Chem. 267:2494-2501). Transferases
Transferases are enzymes that catalyze the transfer of molecular groups. The reaction may involve an oxidation, reduction, or cleavage of covalent bonds, and is often specific to a substrate or to particular sites on a type of substrate. Transferases participate in reactions essential to such 5 functions as synthesis and degradation of cell components, regulation of cell functions including cell signaling, cell proliferation, inflamation, apoptosis, secretion and excretion. Transferases are involved in key steps in disease processes involving these functions. Transferases are frequently classified according to the type of group transferred. For example, methyl transferases transfer one- carbon methyl groups, amino transferases transfer nitrogenous amino groups, and similarly o denominated enzymes transfer aldehyde or ketone, acyl, glycosyl, alkyl or aryl, isoprenyl, saccharyl, phosphorous-containing, sulfur-containing, or selenium-containing groups, as well as small enzymatic groups such as Coenzyme A.
Acyl transferases include peroxisomal carnitine octanoyl transferase, which is involved in the fatty acid beta-oxidation pathway, and mitochondrial carnitine palmitoyl transferases, involved in 5 fatty acid metabolism and transport. Choline O-acetyl transferase catalyzes the biosynthesis of the neurotransmitter acetylcholine.
Amino transferases play key roles in protein synthesis and degradation, and they contribute to other processes as well. For example, the amino transferase 5-aminolevulinic acid synthase catalyzes the addition of succinyl-CoA to glycine, the first step in heme biosynthesis. Other amino transferases 5 participate in pathways important for neurological function and metabolism. For example, glutamine- phenylpyruvate amino transferase, also known as glutamine transaminase K (GTK), catalyzes several reactions with a pyridoxal phosphate cofactor. GTK catalyzes the reversible conversion of L- glutamine and phenylpyruvate to 2-oxoglutaramate and L-phenylalanine. Other amino acid substrates for GTK include L-methionine, L-histidine, and L-tyrosine. GTK also catalyzes the conversion of o kynurenine to kynurenic acid, a tryptophan metabolite that is an antagonist of the N-methyl-D- aspartate (NMD A) receptor in the brain and may exert a neuromodulatory function. Alteration of the kynurenine metabolic pathway may be associated with several neurological disorders. GTK also plays a role in the metabolism of halogenated xenobiotics conjugated to glutathione, leading to nephrotoxicity in rats and neurotoxicity in humans. GTK is expressed in kidney, liver, and brain. 5 Both human and rat GTKs contain a putative pyridoxal phosphate binding site (ExPASy ENZYME: EC 2.6.1.64; Perry, S.J. et al. (1993) Mol. Pharmacol. 43:660-665; Perry, S. et al. (1995) FEBS Lett. 360:277-280; and Alberati-Giani, D. et al. (1995) J. Neurochem. 64:1448-1455). A second amino transferase associated with this pathway is kynurenine/α-aminoadipate amino transferase (AadAT). AadAT catalyzes the reversible conversion of α-aminoadipate and α-ketoglutarate to α-ketoadipate o and L-glutamate during lysine metabolism. AadAT also catalyzes the transamination of kynurenine to kynurenic acid. A cytosolic AadAT is expressed in rat kidney, liver, and brain (Nakatani, Y. et al. (1970) Biochim. Biophys. Acta 198:219-228; Buchli, R. et al. (1995) J. Biol. Chem. 270:29330- 29335).
Glycosyl transferases include the mammalian UDP-glucouronosyl transferases, a family of 5 membrane-bound microsomal enzymes catalyzing the transfer of glucouronic acid to lipophilic substrates in reactions that play important roles in detoxification and excretion of drugs, carcinogens, and other foreign substances. Another mammalian glycosyl transferase, mammalian UDP-galactose- ceramide galactosyl transferase, catalyzes the transfer of galactose to ceramide in the synthesis of galactocerebrosides in myelin membranes of the nervous system. The UDP-glycosyl transferases o share a conserved signature domain of about 50 amino acid residues (PROSITE: PDOC00359, http://expasy.hcuge.ch/sprot/prosite.html).
Methyl transferases are involved in a variety of pharmacologically important processes. Nicotinamide N-methyl transferase catalyzes the N-methylation of nicotinamides and other pyridines, an important step in the cellular handling of drugs and other foreign compounds. 5 Phenylethanolamine N-methyl transferase catalyzes the conversion of noradrenalin to adrenalin. 6-0- methylguanine-DNA methyl transferase reverses DNA methylation, an important step in carcinogenesis. Uroporphyrin-III C-methyl transferase, which catalyzes the transfer of two methyl groups from S-adenosyl-L-methionine to uroporphyrinogen III, is the first specific enzyme in the biosynthesis of cobalamin, a dietary enzyme whose uptake is deficient in pernicious anemia. Protein- 5 arginine methyl transferases catalyze the posttranslational methylation of arginine residues in proteins, resulting in the mono- and dimethylation of arginine on the guanidino group. Substrates include histones, myelin basic protein, and heterogeneous nuclear ribonucleoproteins involved in mRNA processing, splicing, and transport. Protein-arginine methyl transferase interacts with proteins upregulated by mitogens, with proteins involved in chronic lymphocytic leukemia, and with 0 interferon, suggesting an important role for methylation in cytokine receptor signaling (Lin, W.-J. et al. (1996) J. Biol. Chem. 271:15034-15044; Abramovich, C. et al. (1997) EMBO J. 16:260-266; and Scott, H.S. et al. (1998) Genomics 48:330-340).
Phosphotransferases catalyze the transfer of high-energy phosphate groups and are important in energy-requiring and -releasing reactions. The metabolic enzyme creatine kinase catalyzes the 5 reversible phosphate transfer between creatine/creatine phosphate and ATP/ADP. Glycocyamine kinase catalyzes phosphate transfer from ATP to guanidoacetate, and arginine kinase catalyzes phosphate transfer from ATP to arginine. A cysteine-containing active site is conserved in this family (PROSITE: PDOC00103).
Prenyl transferases are heterodimers, consisting of an alpha and a beta subunit, that catalyze 0 the transfer of an isoprenyl group. An example of a prenyl transferase is the mammalian protein farnesyl transferase. The alpha subunit of farnesyl transferase consists of 5 repeats of 34 amino acids each, with each repeat containing an invariant tryptophan (PROSITE: PDOC00703).
Saccharyl transferases are glycating enzymes involved in a variety of metabolic processes. Oligosacchryl transferase-48, for example, is a receptor for advanced glycation endproducts. 5 Accumulation of these endproducts is observed in vascular complications of diabetes, macrovascular disease, renal insufficiency, and Alzheimer's disease (Thornalley, P.J. (1998) Cell Mol. Biol. (Noisy- Le-Grand) 44:1013-1023).
Coenzyme A (Co A) transferase catalyzes the transfer of Co A between two carboxylic acids. Succinyl CoA:3-oxoacid CoA transferase, for example, transfers CoA from succinyl-CoA to a o recipient such as acetoacetate. Acetoacetate is essential to the metabolism of ketone bodies, which accumulate in tissues affected by metabolic disorders such as diabetes (PROSITE: PDOC00980). Hydrolases
Hydrolysis is the breaking of a covalent bond in a substrate by introduction of a molecule of water. The reaction involves a nucleophilic attack by the water molecule's oxygen atom on a target 5 bond in the substrate. The water molecule is split across the target bond, breaking the bond and generating two product molecules. Hydrolases participate in reactions essential to such functions as synthesis and degradation of cell components, and for regulation of cell functions including cell signaling, cell proliferation, inflamation, apoptosis, secretion and excretion. Hydrolases are involved in key steps in disease processes involving these functions. Hydrolytic enzymes, or hydrolases, may 5 be grouped by substrate specificity into classes including phosphatases, peptidases, lysophospholipases, phosphodiesterases, glycosidases, and glyoxalases.
Phosphatases hydrolytically remove phosphate groups from proteins, an energy-providing step that regulates many cellular processes, including intracellular signaling pathways that in turn control cell growth and differentiation, cell-cell contact, the cell cycle, and oncogenesis. l o Lysophospholipases (LPLs) regulate intracellular lipids by catalyzing the hydrolysis of ester bonds to remove an acyl group, a key step in lipid degradation. Small LPL isoforms, approximately 15-30 kD, function as hydrolases; larger isoforms function both as hydrolases and transacylases. A particular substrate for LPLs, lysophosphatidylcholine, causes lysis of cell membranes. LPL activity is regulated by signaling molecules important in numerous pathways, including the inflammatory
15 response.
Peptidases, also called proteases, cleave peptide bonds that form the backbone of peptide or protein chains. Proteolytic processing is essential to cell growth, differentiation, remodeling, and homeostasis as well as inflammation and immune response. Since typical protein half -lives range from hours to a few days, peptidases are continually cleaving precursor proteins to their active form,
20 removing signal sequences from targeted proteins, and degrading aged or defective proteins.
Peptidases function in bacterial, parasitic, and viral invasion and replication within a host. Examples of peptidases include trypsin and chymotrypsin (components of the complement cascade and the blood-clotting cascade) lysosomal cathepsins, calpains, pepsin, renin, and chymosin (Beynon, R.J. and J.S. Bond (1994) Proteolytic Enzymes: A Practical Approach, Oxford University Press, New
25 York NY, pp. 1-5).
The phosphodiesterases catalyze the hydrolysis of one of the two ester bonds in a phosphodiester compound. Phosphodiesterases are therefore crucial to a variety of cellular processes. Phosphodiesterases include DNA and RNA endo- and exo-nucleases, which are essential to cell growth and replication as well as protein synthesis. Another phosphodiesterase is acid
3 o sphingomyeUnase, which hydrolyzes the membrane phosphoUpid sphingomyelin to ceramide and phosphorylcholine. Phosphorylcholine is used in the synthesis of phosphatidylcholine, which is involved in numerous intracellular signaling pathways. Ceramide is an essential precursor for the generation of gangUosides, membrane lipids found in high concentration in neural tissue. Defective acid sphingomyeUnase phosphodiesterase leads to a build-up of sphingomyelin molecules in
35 lysosomes, resulting in Niemann-Pick disease. Glycosidases catalyze the cleavage of hemiacetyl bonds of glycosides, which are compounds that contain one or more sugar. Mammalian lactase-phlorizin hydrolase, for example, is an intestinal enzyme that splits lactose. Mammalian beta-galactosidase removes the terminal galactose from gangliosides, glycoproteins, and glycosaminoglycans, and deficiency of this enzyme is associated 5 with a gangliosidosis known as Morquio disease type B. Vertebrate lysosomal alpha-glucosidase, which hydrolyzes glycogen, maltose, and isomaltose, and vertebrate intestinal sucrase-isomaltase, which hydrolyzes sucrose, maltose, and isomaltose, are widely distributed members of this family with highly conserved sequences at their active sites.
The glyoxylase system is involved in gluconeogenesis, the production of glucose from o storage compounds in the body. It consists of glyoxylase I, which catalyzes the formation of S-D- lactoylglutathione from methyglyoxal, a side product of triose-phosphate energy metaboUsm, and glyoxylase II, which hydrolyzes S-D-lactoylglutathione to D-lactic acid and reduced glutathione. Glyoxylases are involved in hyperglycemia, non-insuUn-dependent diabetes mellitus, the detoxification of bacterial toxins, and in the control of cell proliferation and microtubule assembly. 5 Lvases
Lyases are a class of enzymes that catalyze the cleavage of C-C, C-O, C-N, C-S, C-(halide), P-0 or other bonds without hydrolysis or oxidation to form two molecules, at least one of which contains a double bond (Stryer, L. (1995) Biochemistry W.H. Freeman and Co. New York, NY p.620). Lyases are critical components of cellular biochemistry with roles in metabolic energy 0 production including fatty acid metabolism, as well as other diverse enzymatic processes. Further classification of lyases reflects the type of bond cleaved as well as the nature of the cleaved group.
The group of C-C lyases include carboxyl-lyases (decarboxylases), aldehyde-lyases (aldolases), oxo-acid-lyases and others. The C-O lyase group includes hydro-lyases, lyases acting on polysaccharides and other lyases. The C-N lyase group includes ammonia-lyases, amidine-lyases, 5 amine-lyases (deaminases) and other lyases.
Proper regulation of lyases is critical to normal physiology. For example, mutation induced deficiencies in the uroporphyrinogen decarboxylase can lead to photosensitive cutaneous lesions in the genetically-Unked disorder familial porphyria cutanea tarda (Mendez, M. et al. (1998) Am. J. Genet. 63: 1363- 1375). It has also been shown that adenosine deaminase (ADA) deficiency stems o from genetic mutations in the ADA gene, resulting in the disorder severe combined immunodeficiency disease (SCID) (Hershfield, M.S. (1998) Semin. Hematol. 35:291-298). Isomerases
Isomerases are a class of enzymes that catalyze geometric or structural changes within a molecule to form a single product. This class includes racemases and epimerases, cis-trans- 5 isomerases, intramolecular oxidoreductases, intramolecular transferases (mutases) and intramolecular lyases. Isomerases are critical components of cellular biochemistry with roles in metabolic energy production including glycolysis, as well as other diverse enzymatic processes (Stryer, L. (1995) Biochemistry, W.H. Freeman and Co., New York NY, pp.483-507).
Racemases are a subset of isomerases that catalyze inversion of a molecules configuration 5 around the asymmetric carbon atom in a substrate having a single center of asymmetry, thereby interconverting two racemers. Epimerases are another subset of isomerases that catalyze inversion of configuration around an asymmetric carbon atom in a substrate with more than one center of symmetry, thereby interconverting two epimers. Racemases and epimerases can act on amino acids and derivatives, hydroxy acids and derivatives, as well as carbohydrates and derivatives. The l o interconversion of UDP-galactose and UDP-glucose is catalyzed by UDP-galactose-4' -epimerase. Proper regulation and function of this epimerase is essential to the synthesis of glycoproteins and glycolipids. Elevated blood galactose levels have been correlated with UDP-galactose-4' -epimerase deficiency in screening programs of infants (Gitzelmann, R. (1972) Helv. Paediat. Acta 27:125-130). Oxidoreductases can be isomerases as well. Oxidoreductases catalyze the reversible transfer
15 of electrons from a substrate that becomes oxidized to a substrate that becomes reduced. This class of enzymes includes dehydrogenases, hydroxylases, oxidases, oxygenases, peroxidases, and reductases. Proper maintenance of oxidoreductase levels is physiologically important. For example, genetically- linked deficiencies in lipoamide dehydrogenase can result in lactic acidosis (Robinson, B.H. et al. (1977) Pediat. Res. 11:1198-1202).
20 Another subgroup of isomerases are the transferases (or mutases). Transferases transfer a chemical group from one compound (the donor) to another compound (the acceptor). The types of groups transferred by these enzymes include acyl groups, amino groups, phosphate groups (phosphotransferases or phosphomutases), and others. The transferase carnitine palmitoyltransferase is an important component of fatty acid metabolism. Genetically-Unked deficiencies in this
25 transferase can lead to myopathy (Scriver, CR. et al. (1995) The Metabolic and Molecular Basis of Inherited Disease. McGraw-Hill, New York NY, pp.1501-1533).
Yet another subgroup of isomerases are the topoisomersases. Topoisomerases are enzymes that affect the topological state of DNA. For example, defects in topoisomerases or their regulation can affect normal physiology. Reduced levels of topoisomerase II have been correlated with some of
30 the DNA processing defects associated with the disorder ataxia-telangiectasia (Singh, S.P. et al. (1988) Nucleic Acids Res. 16:3919-3929). Li gases
Ligases catalyze the formation of a bond between two substrate molecules. The process involves the hydrolysis of a pyrophosphate bond in ATP or a similar energy donor. Ligases are
35 classified based on the nature of the type of bond they form, which can include carbon-oxygen, carbon-sulfur, carbon-nitrogen, carbon-carbon and phosphoric ester bonds.
Ligases forming carbon-oxygen bonds include the aminoacyl-transfer RNA (tRNA) synthetases which are important RNA-associated enzymes with roles in translation. Protein biosynthesis depends on each amino acid forming a Unkage with the appropriate tRNA. The 5 aminoacyl-tRNA synthetases are responsible for the activation and correct attachment of an amino acid with its cognate tRNA. The 20 aminoacyl-tRNA synthetase enzymes can be divided into two structural classes, and each class is characterized by a distinctive topology of the catalytic domain. Class I enzymes contain a catalytic domain based on the nucleotide-binding Rossman fold. Class II enzymes contain a central catalytic domain, which consists of a seven-stranded antiparallel β-sheet 0 motif, as well as N- and C- terminal regulatory domains. Class II enzymes are separated into two groups based on the heterodimeric or homodimeric structure of the enzyme; the latter group is further subdivided by the structure of the N- and C-terminal regulatory domains (Hartlein, M. and S. Cusack (1995) J. Mol. Evol. 40:519-530). Autoantibodies against aminoacyl-tRNAs are generated by patients with dermatomyositis and polymyositis, and correlate strongly with complicating interstitial 5 lung disease (ILD). These antibodies appear to be generated in response to viral infection, and coxsackie virus has been used to induce experimental viral myositis in animals.
Ligases forming carbon-sulfur bonds (Acid-thiol ligases) mediate a large number of cellular biosynthetic intermediary metabolism processes involve intermolecular transfer of carbon atom-containing substrates (carbon substrates). Examples of such reactions include the tricarboxylic o acid cycle, synthesis of fatty acids and long-chain phospholipids, synthesis of alcohols and aldehydes, synthesis of intermediary metabolites, and reactions involved in the amino acid degradation pathways. Some of these reactions require input of energy, usually in the form of conversion of ATP to either ADP or AMP and pyrophosphate.
In many cases, a carbon substrate is derived from a small molecule containing at least two 5 carbon atoms. The carbon substrate is often covalently bound to a larger molecule which acts as a carbon substrate carrier molecule within the cell. In the biosynthetic mechanisms described above, the carrier molecule is coenzyme A. Coenzyme A (CoA) is structurally related to derivatives of the nucleotide ADP and consists of 4'-phosphopantetheine linked via a phosphodiester bond to the alpha phosphate group of adenosine 3',5'-bisphosphate. The terminal thiol group of 4'-phosphopantetheine o acts as the site for carbon substrate bond formation. The predominant carbon substrates which utilize
CoA as a carrier molecule during biosynthesis and intermediary metabolism in the cell are acetyl, succinyl, and propionyl moieties, collectively referred to as acyl groups. Other carbon substrates include enoyl lipid, which acts as a fatty acid oxidation intermediate, and carnitine, which acts as an acetyl-CoA flux regulator/ mitochondrial acyl group transfer protein. Acyl-CoA and acetyl-CoA are 5 synthesized in the cell by acyl-CoA synthetase and acetyl-CoA synthetase, respectively. Activation of fatty acids is mediated by at least three forms of acyl-CoA synthetase activity: i) acetyl-CoA synthetase, which activates acetate and several other low molecular weight carboxylic acids and is found in muscle mitochondria and the cytosol of other tissues; u) medium-chain acyl-CoA synthetase, which activates fatty acids containing between four and eleven carbon atoms 5 (predominantly from dietary sources), and is present only in liver mitochondria; and in) acyl CoA synthetase, which is specific for long chain fatty acids with between six and twenty carbon atoms, and is found in microsomes and the mitochondria. Proteins associated with acyl-CoA synthetase activity have been identified from many sources including bacteria, yeast, plants, mouse, and man. The activity of acyl-CoA synthetase may be modulated by phosphorylation of the enzyme by o cAMP-dependent protein kinase.
Ligases forming carbon-nitrogen bonds include amide synthases such as glutamine synthetase (glutamate-ammonia ligase) that catalyzes the animation of glutamic acid to glutamine by ammonia using the energy of ATP hydrolysis. Glutamine is the primary source for the amino group in various amide transfer reactions involved in de novo pyrimidine nucleotide synthesis and in purine and 5 pyrimidine ribonucleotide interconversions. Overexpression of glutamine synthetase has been observed in primary liver cancer (Christa, L. et al. (1994) Gastroent. 106:1312-1320).
Acid-amino-acid ligases (peptide synthases) are represented by the ubiquitin proteases which are associated with the ubiquitin conjugation system (UCS), a major pathway for the degradation of cellular proteins in eukaryotic cells and some bacteria. The UCS mediates the elimination of o abnormal proteins and regulates the half -lives of important regulatory proteins that control cellular processes such as gene transcription and cell cycle progression. In the UCS pathway, proteins targeted for degradation are conjugated to a ubiquitin (Ub), a small heat stable protein. Ub is first activated by a ubiquitin-activating enzyme (El), and then transferred to one of several Ub- conjugating enzymes (E2). E2 then links the Ub molecule through its C-terminal glycine to an 5 internal lysine (acceptor lysine) of a target protein. The ubiquitinated protein is then recognized and degraded by proteasome, a large, multisubunit proteolytic enzyme complex, and ubiquitin is released for reutilization by ubiquitin protease. The UCS is implicated in the degradation of mitotic cyclic kinases, oncoproteins, tumor suppressor genes such as p53, viral proteins, cell surface receptors associated with signal transduction, transcriptional regulators, and mutated or damaged proteins o (Ciechanover, A. (1994) Cell 79: 13-21). A murine proto-oncogene, Unp, encodes a nuclear ubiquitin protease whose overexpression leads to oncogenic transformation of NIH3T3 cells, and the human homolog of this gene is consistently elevated in small cell tumors and adenocarcinomas of the lung (Gray, D.A. (1995) Oncogene 10:2179-2183).
Cyclo-ligases and other carbon-nitrogen ligases comprise various enzymes and enzyme 5 complexes that participate in the de novo pathways to purine and pyrimidine biosynthesis. Because these pathways are critical to the synthesis of nucleotides for replication of both RNA and DNA, many of these enzymes have been the targets of clinical agents for the treatment of cell proliferative disorders such as cancer and infectious diseases.
Purine biosynthesis occurs de novo from the amino acids glycine and glutamine, and other 5 small molecules. Three of the key reactions in this process are catalyzed by a trifunctional enzyme composed of glycinamide-ribonucleotide synthetase (GARS), aminoimidazole ribonucleotide synthetase (AIRS), and glycinamide ribonucleotide transformylase (GART). Together these three enzymes combine ribosylamine phosphate with glycine to yield phosphoribosyl aminoimidazole, a precursor to both adenylate and guanylate nucleotides. This trifunctional protein has been impUcated 0 in the pathology of Downs syndrome (Aimi, J. et al. (1990) Nucleic Acid Res. 18:6665-6672).
Adenylosuccinate synthetase catalyzes a later step in purine biosynthesis that converts inosinic acid to adenylosuccinate, a key step on the path to ATP synthesis. This enzyme is also similar to another carbon-nitrogen ligase, argininosuccinate synthetase, that catalyzes a similar reaction in the urea cycle (Powell, S.M. et al. (1992) FEBS Lett. 303:4-10). 5 Like the de novo biosynthesis of purines, de novo synthesis of the pyrimidine nucleotides uridylate and cytidylate also arises from a common precursor, in this instance the nucleotide orotidylate derived from orotate and phosphoribosyl pyrophosphate (PPRP). Again a trifunctional enzyme comprising three carbon-nitrogen ligases plays a key role in the process. In this case the enzymes aspartate transcarbamylase (ATCase), carbamyl phosphate synthetase II, and dihydroorotase o (DHOase) are encoded by a single gene called CAD. Together these three enzymes combine the initial reactants in pyrimidine biosynthesis, glutamine, C02 and ATP to form dihydroorotate, the precursor to orotate and orotidylate (Iwahana, H. et al. (1996) Biochem. Biophys. Res. Commun. 219:249-255). Further steps then lead to the synthesis of uridine nucleotides from orotidylate. Cytidine nucleotides are derived from uridine-5'-triphosphate (UTP) by the amidation of UTP using 5 glutamine as the amino donor and the enzyme CTP synthetase. Regulatory mutations in the human CTP synthetase are believed to confer multi-drug resistance to agents widely used in cancer therapy (Yamauchi, M. et al. (1990) EMBO J. 9:2095-2099).
Ligases forming carbon-carbon bonds include the carboxylases acetyl-CoA carboxylase and pyruvate carboxylase. Acetyl-CoA carboxylase catalyzes the carboxylation of acetyl-CoA from C02 o and H20 using the energy of ATP hydrolysis. Acetyl-CoA carboxylase is the rate-limiting step in the biogenesis of long-chain fatty acids. Two isoforms of acetyl-CoA carboxylase, types I and types II, are expressed in human in a tissue-specific manner (Ha, J. et al. (1994) Eur. J. Biochem. 219:297- 306). Pyruvate carboxylase is a nuclear-encoded mitochondrial enzyme that catalyzes the conversion of pyruvate to oxaloacetate, a key intermediate in the citric acid cycle. 5 Ligases forming phosphoric ester bonds include the DNA ligases involved in both DNA replication and repair. DNA ligases seal phosphodiester bonds between two adjacent nucleotides in a DNA chain using the energy from ATP hydrolysis to first activate the free 5 '-phosphate of one nucleotide and then react it with the 3' -OH group of the adjacent nucleotide. This reseating reaction is used in both DNA replication to join small DNA fragments called Okazaki fragments that are transiently formed in the process of replicating new DNA, and in DNA repair. DNA repair is the process by which accidental base changes, such as those produced by oxidative damage, hydrolytic attack, or uncontrolled methylation of DNA, are corrected before replication or transcription of the DNA can occur. Bloom's syndrome is an inherited human disease in which individuals are partially deficient in DNA Ugation and consequently have an increased incidence of cancer (Alberts, B. et al. (1994) The Molecular Biology of the Cell, Garland Publishing Inc., New York NY, p. 247).
Molecules Associated with Growth and Development
SEQ ID NO:69, SEQ ID NO:70, and SEQ ID NO:71 encode, for example, molecules associated with growth and development. Human growth and development requires the spatial and temporal regulation of cell differentiation, cell proliferation, and apoptosis. These processes coordinately control reproduction, aging, embryogenesis, morphogenesis, organogenesis, and tissue repair and maintenance. At the cellular level, growth and development is governed by the cell's decision to enter into or exit from the cell division cycle and by the cell's commitment to a terminally differentiated state. These decisions are made by the cell in response to extracellular signals and other environmental cues it receives. The following discussion focuses on the molecular mechanisms of cell division, reproduction, cell differentiation and proUferation, apoptosis, and aging. Cell Division
Cell division is the fundamental process by which all living things grow and reproduce. In unicellular organisms such as yeast and bacteria, each cell division doubles the number of organisms, while in multicellular species many rounds of cell division are required to replace cells lost by wear or by programmed cell death, and for cell differentiation to produce a new tissue or organ. Details of the cell division cycle may vary, but the basic process consists of three principle events. The first event, interphase, involves preparations for cell division, repUcation of the DNA, and production of essential proteins. In the second event, mitosis, the nuclear material is divided and separates to opposite sides of the cell. The final event, cytokinesis, is division and fission of the cell cytoplasm. The sequence and timing of cell cycle transitions is under the control of the cell cycle regulation system which controls the process by positive or negative regulatory circuits at various check points.
Regulated progression of the cell cycle depends on the integration of growth control pathways with the basic cell cycle machinery. Cell cycle regulators have been identified by selecting for human and yeast cDNAs that block or activate cell cycle arrest signals in the yeast mating pheromone pathway when they are overexpressed. Known regulators include human CPR (cell cycle progression restoration) genes, such as CPR8 and CPR2, and yeast CDC (cell division control) genes, including CDC91, that block the arrest signals. The CPR genes express a variety of proteins including cycUns, tumor suppressor binding proteins, chaperones, transcription factors, translation factors, and RNA-binding proteins (Edwards, M.C et al.(1997) Genetics 147:1063-1076).
Several cell cycle transitions, including the entry and exit of a cell from mitosis, are dependent upon the activation and inhibition of cycUn-dependent kinases (Cdks). The Cdks are composed of a kinase subunit, Cdk, and an activating subunit, cycUn, in a complex that is subject to many levels of regulation. There appears to be a single Cdk in Saccharomyces cerevisiae and Saccharomyces pombe whereas mammals have a variety of speciaUzed Cdks. CycUns act by binding to and activating cycUn-dependent protein kinases which then phosphorylate and activate selected proteins involved in the mitotic process. The Cdk-cycUn complex is both positively and negatively regulated by phosphorylation, and by targeted degradation involving molecules such as CDC4 and CDC53. In addition, Cdks are further regulated by binding to inhibitors and other proteins such as Sucl that modify their specificity or accessibiUty to regulators (Patra, D. and W.G. Dunphy (1996) Genes Dev. 10:1503-1515; and Mathias, N. et al. (1996) Mol. Cell Biol. 16:6634-6643). Reproduction The male and female reproductive systems are complex and involve many aspects of growth and development. The anatomy and physiology of the male and female reproductive systems are reviewed in (Guyton, A.C. (1991) Textbook of Medical Physiology. W.B. Saunders Co., Philadelphia PA, pp. 899-928).
The male reproductive system includes the process of spermatogenesis, in which the sperm are formed, and male reproductive functions are regulated by various hormones and their effects on accessory sexual organs, cellular metabolism, growth, and other bodily functions.
Spermatogenesis begins at puberty as a result of stimulation by gonadotropic hormones released from the anterior pituitary. Immature sperm (spermatogonia) undergo several mitotic cell divisions before undergoing meiosis and full maturation. The testes secrete several male sex hormones, the most abundant being testosterone, that is essential for growth and division of the immature sperm, and for the masculine characteristics of the male body. Three other male sex hormones, gonadotropin- releasing hormone (GnRH), luteinizing hormone (LH), and follicle-stimulating hormone (FSH) control sexual function.
The uterus, ovaries, fallopian tubes, vagina, and breasts comprise the female reproductive system. The ovaries and uterus are the source of ova and the location of fetal development, respectively. The fallopian tubes and vagina are accessory organs attached to the top and bottom of the uterus, respectively. Both the uterus and ovaries have additional roles in the development and loss of reproductive capabiUty during a female' s Ufetime. The primary role of the breasts is lactation. 5 Multiple endocrine signals from the ovaries, uterus, pituitary, hypothalamus, adrenal glands, and other tissues coordinate reproduction and lactation. These signals vary during the monthly menstruation cycle and during the female's Ufetime. Similarly, the sensitivity of reproductive organs to these endocrine signals varies during the female's lifetime.
A combination of positive and negative feedback to the ovaries, pituitary and hypothalamus o glands controls physiologic changes during the monthly ovulation and endometrial cycles. The anterior pituitary secretes two major gonadotropin hormones, folUcle-stimulating hormone (FSH) and luteinizing hormone (LH), regulated by negative feedback of steroids, most notably by ovarian estradiol. If fertiUzation does not occur, estrogen and progesterone levels decrease. This sudden reduction of the ovarian hormones leads to menstruation, the desquamation of the endometrium. 5 Hormones further govern all the steps of pregnancy, parturition, lactation, and menopause.
During pregnancy large quantities of human chorionic gonadotropin (hCG), estrogens, progesterone, and human chorionic somatomammotropin (hCS) are formed by the placenta. hCG, a glycoprotein similar to luteinizing hormone, stimulates the corpus luteum to continue producing more progesterone and estrogens, rather than to involute as occurs if the ovum is not fertiUzed. hCS is similar to growth o hormone and is crucial for fetal nutrition.
The female breast also matures during pregnancy. Large amounts of estrogen secreted by the placenta trigger growth and branching of the breast milk ductal system while lactation is initiated by the secretion of prolactin by the pituitary gland.
Parturition involves several hormonal changes that increase uterine contractility toward the end 5 of pregnancy, as follows. The levels of estrogens increase more than those of progesterone. Oxytocin is secreted by the neurohypophysis. Concomitantly, uterine sensitivity to oxytocin increases. The fetus itself secretes oxytocin, cortisol (from adrenal glands), and prostaglandins.
Menopause occurs when most of the ovarian follicles have degenerated. The ovary then produces less estradiol, reducing the negative feedback on the pituitary and hypothalamus glands. o Mean levels of circulating FSH and LH increase, even as ovulatory cycles continue. Therefore, the ovary is less responsive to gonadotropins, and there is an increase in the time between menstrual cycles. Consequently, menstrual bleeding ceases and reproductive capabiUty ends. Cell Differentiation and ProUferation
Tissue growth involves complex and ordered patterns of cell proliferation, cell differentiation, and apoptosis. Cell proliferation must be regulated to maintain both the number of cells and their spatial organization. This regulation depends upon the appropriate expression of proteins which control cell cycle progression in response to extracellular signals, such as growth factors and other mitogens, and intracellular cues, such as DNA damage or nutrient starvation. Molecules which directly or indirectly modulate cell cycle progression fall into several categories, including growth factors and their receptors, second messenger and signal transduction proteins, oncogene products, tumor-suppressor proteins, and mitosis-promoting factors.
Growth factors were originally described as serum factors required to promote cell proliferation. Most growth factors are large, secreted polypeptides that act on cells in their local environment. Growth factors bind to and activate specific cell surface receptors and initiate intracellular signal transduction cascades. Many growth factor receptors are classified as receptor tyrosine kinases which undergo autophosphorylation upon ligand binding. Autophosphorylation enables the receptor to interact with signal transduction proteins characterized by the presence of SH2 or SH3 domains (Src homology regions 2 or 3). These proteins then modulate the activity state of small G-proteins, such as Ras, Rab, and Rho, along with GTPase activating proteins (GAPs), guanine nucleotide releasing proteins (GNRPs), and other guanine nucleotide exchange factors. Small G proteins act as molecular switches that activate other downstream events, such as mitogen-activated protein kinase (MAP kinase) cascades. MAP kinases ultimately activate transcription of mitosis- promoting genes. In addition to growth factors, small signaUng peptides and hormones also influence cell proUferation. These molecules bind primarily to another class of receptor, the trimeric G-protein coupled receptor (GPCR), found predominantly on the surface of immune, neuronal and neuroendocrine cells. Upon ligand binding, the GPCR activates a trimeric G protein which in turn triggers increased levels of intracellular second messengers such as phosphoUpase C, Ca2+, and cycUc AMP. Most GPCR-mediated signaUng pathways indirectly promote cell proliferation by causing the secretion or breakdown of other signaUng molecules that have direct mitogenic effects. These signaling cascades often involve activation of kinases and phosphatases. Some growth factors, such as some members of the transforming growth factor beta (TGF-β) family, act on some cells to stimulate cell proliferation and on other cells to inhibit it. Growth factors may also stimulate a cell at one concentration and inhibit the same cell at another concentration. Most growth factors also have a multitude of other actions besides the regulation of cell growth and division: they can control the proUferation, survival, differentiation, migration, or function of cells depending on the circumstance. For example, the tumor necrosis factor/nerve growth factor (TNF/NGF) family can activate or inhibit cell death, as well as regulate proUferation and differentiation. The cell response depends on the type of cell, its stage of differentiation and transformation status, which surface receptors are stimulated, and the types of stimuU acting on the cell (Smith, A. et al. (1994) Cell 76:959-962; and Nocentini, G. et al. (1997) Proc. Natl. Acad. Sci. USA 94:6216-6221).
Neighboring cells in a tissue compete for growth factors, and when provided with "unUmited" quantities in a perfused system will grow to even higher cell densities before reaching density-dependent inhibition of cell division. Cells often demonstrate an anchorage dependence of cell division as well. This anchorage dependence may be associated with the formation of focal contacts Unking the cytoskeleton with the extracellular matrix (ECM). The expression of ECM components can be stimulated by growth factors. For example, TGF-β stimulates fibroblasts to produce a variety of ECM proteins, including fibronectin, collagen, and tenascin (Pearson, CA. et al. (1988) EMBO J. 7:2677- 2981). In fact, for some cell types specific ECM molecules, such as laminin or fibronectin, may act as growth factors. Tenascin-C and -R, expressed in developing and lesioned neural tissue, provide stimulatory/anti-adhesive or inhibitory properties, respectively, for axonal growth (Faissner, A. (1997) Cell Tissue Res. 290:331-341). Cancers are associated with the activation of oncogenes which are derived from normal cellular genes. These oncogenes encode oncoproteins which convert normal cells into maUgnant cells. Some oncoproteins are mutant isoforms of the normal protein, and other oncoproteins are abnormally expressed with respect to location or amount of expression. The latter category of oncoprotein causes cancer by altering transcriptional control of cell proUferation. Five classes of oncoproteins are known to affect cell cycle controls. These classes include growth factors, growth factor receptors, intracellular signal transducers, nuclear transcription factors, and cell-cycle control proteins. Viral oncogenes are integrated into the human genome after infection of human cells by certain viruses. Examples of viral oncogenes include v-src, v-abl, and v-fps.
Many oncogenes have been identified and characterized. These include sis, erbA, erbB, her-2, mutated Gs, src, abl, ras, crk, jun, fos, myc, and mutated tumor-suppressor genes such as RB, p53, mdm2, Cipl, pi 6, and cyclin D. Transformation of normal genes to oncogenes may also occur by chromosomal translocation. The Philadelphia chromosome, characteristic of chronic myeloid leukemia and a subset of acute lymphoblastic leukemias, results from a reciprocal translocation between chromosomes 9 and 22 that moves a truncated portion of the proto-oncogene c-abl to the breakpoint cluster region (bcr) on chromosome 22.
Tumor-suppressor genes are involved in regulating cell proUferation. Mutations which cause reduced or loss of function in tumor-suppressor genes result in uncontrolled cell proUferation. For example, the retinoblastoma gene product (RB), in a non-phosphorylated state, binds several early- response genes and suppresses their transcription, thus blocking cell division. Phosphorylation of RB causes it to dissociate from the genes, releasing the suppression, and allowing cell division to proceed. Apoptosis
Apoptosis is the genetically controlled process by which unneeded or defective cells undergo programmed cell death. Selective eUmination of cells is as important for morphogenesis and tissue 5 remodeUng as is cell proliferation and differentiation. Lack of apoptosis may result in hyperplasia and other disorders associated with increased cell proUferation. Apoptosis is also a critical component of the immune response. Immune cells such as cytotoxic T-cells and natural killer cells prevent the spread of disease by inducing apoptosis in tumor cells and virus-infected cells. In addition, immune cells that fail to distinguish self molecules from foreign molecules must be eUminated by apoptosis to avoid an o autoimmune response.
Apoptotic cells undergo distinct morphological changes. Hallmarks of apoptosis include cell shrinkage, nuclear and cytoplasmic condensation, and alterations in plasma membrane topology. Biochemically, apoptotic cells are characterized by increased intracellular calcium concentration, fragmentation of chromosomal DNA, and expression of novel cell surface components. 5 The molecular mechanisms of apoptosis are highly conserved, and many of the key protein regulators and effectors of apoptosis have been identified. Apoptosis generally proceeds in response to a signal which is transduced intracellularly and results in altered patterns of gene expression and protein activity. Signaling molecules such as hormones and cytokines are known both to stimulate and to inhibit apoptosis through interactions with cell surface receptors. Transcription factors also play an o important role in the onset of apoptosis. A number of downstream effector molecules, particularly proteases such as the cysteine proteases called caspases, have been implicated in the degradation of cellular components and the proteolytic activation of other apoptotic effectors. Aging and Senescence
Studies of the aging process or senescence have shown a number of characteristic cellular and 5 molecular changes (Fauci et al. (1998) Harrison's Principles of Internal Medicine, McGraw-Hill, New
York NY, p.37). These characteristics include increases in chromosome structural abnormaUties, DNA cross-Unking, incidence of single-stranded breaks in DNA, losses in DNA methylation, and degradation of telomere regions. In addition to these DNA changes, post-translational alterations of proteins increase including, deamidation, oxidation, cross-linking, and nonenzymatic glycation. Still further o molecular changes occur in the mitochondria of aging cells through deterioration of structure. These changes eventually contribute to decreased function in every organ of the body.
Biochemical Pathway Molecules
SEQ ID NO:64, SEQ ID NO:65, SEQ ID NO:66, SEQ ID NO:67, and SEQ ID NO:68 encode, for example, biochemical pathway molecules.
Biochemical pathways are responsible for regulating metaboUsm, growth and development, protein secretion and trafficking, environmental responses, and ecological interactions including immune response and response to parasites. DNA replication
Deoxyribonucleic acid (DNA), the genetic material, is found in both the nucleus and mitochondria of human cells. The bulk of human DNA is nuclear, in the form of Unear chromosomes, while mitochondrial DNA is circular. DNA repUcation begins at specific sites called origins of repUcation. Bidirectional synthesis occurs from the origin via two growing forks that move in opposite directions. Replication is semi-conservative, with each daughter duplex containing one old strand and its newly synthesized complementary partner. Proteins involved in DNA repUcation include DNA polymerases, DNA primase, telomerase, DNA heUcase, topoisomerases, DNA Ugases, repUcation factors, and DNA-binding proteins. DNA Recombination and Repair Cells are constantly faced with repUcation errors and environmental assault (such as ultraviolet irradiation) that can produce DNA damage. Damage to DNA consists of any change that modifies the structure of the molecule. Changes to DNA can be divided into two general classes, single base changes and structural distortions. Any damage to DNA can produce a mutation, and the mutation may produce a disorder, such as cancer. Changes in DNA are recognized by repair systems within the cell. These repair systems act to correct the damage and thus prevent any deleterious affects of a mutational event. Repair systems can be divided into three general types, direct repair, excision repair, and retrieval systems. Proteins involved in DNA repair include DNA polymerase, excision repair proteins, excision and cross Unk repair proteins, recombination and repair proteins, RAD51 proteins, and BLN and WRN proteins that are homologs of RecQ heUcase. When the repair systems are eliminated, cells become exceedingly sensitive to environmental mutagens, such as ultraviolet irradiation. Patients with disorders associated with a loss in DNA repair systems often exhibit a high sensitivity to environmental mutagens. Examples of such disorders include xeroderma pigmentosum (XP), Bloom's syndrome (BS), and Werner's syndrome (WS) (Yamagata, K et al. (1998) Proc. Natl. Acad. Sci. USA 95:8733-8738), ataxia telangiectasia, Cockayne's syndrome, and Fanconi's anemia.
Recombination is the process whereby new DNA sequences are generated by the movements of large pieces of DNA. In homologous recombination, which occurs during meiosis and DNA repair, parent DNA duplexes aUgn at regions of sequence similarity, and new DNA molecules form by the breakage and joining of homologous segments. Proteins involved include RAD51 recombinase. In site- specific recombination, two specific but not necessarily homologous DNA sequences are exchanged. In the immune system this process generates a diverse collection of antibody and T cell receptor genes. Proteins involved in site-specific recombination in the immune system include recombination activating genes 1 and 2 (RAG1 and RAG2). A defect in immune system site-specific recombination causes 5 severe combined immunodeficiency disease in mice. RNA Metabolism
Ribonucleic acid (RNA) is a Unear single-stranded polymer of four nucleotides, ATP, CTP, UTP, and GTP. In most organisms, RNA is transcribed as a copy of DNA, the genetic material of the organism. In retroviruses RNA rather than DNA serves as the genetic material. RNA copies of the o genetic material encode proteins or serve various structural, catalytic, or regulatory roles in organisms.
RNA is classified according to its cellular locaUzation and function. Messenger RNAs (mRNAs) encode polypeptides. Ribosomal RNAs (rRNAs) are assembled, along with ribosomal proteins, into ribosomes, which are cytoplasmic particles that translate mRNA into polypeptides. Transfer RNAs (tRNAs) are cytosoUc adaptor molecules that function in mRNA translation by recognizing both an 5 mRNA codon and the amino acid that matches that codon. Heterogeneous nuclear RNAs (hnRNAs) include mRNA precursors and other nuclear RNAs of various sizes. Small nuclear RNAs (snRNAs) are a part of the nuclear sphceosome complex that removes intervening, non-coding sequences (introns) and rejoins exons in pre-mRNAs. RNA Transcription o The transcription process synthesizes an RNA copy of DNA. Proteins involved include multi- subunit RNA polymerases, transcription factors IIA, IIB, IID, HE, IIF, IIH, and ILL Many transcription factors incorporate DNA-binding structural motifs which comprise either α -helices or β- sheets that bind to the major groove of DNA. Four well-characterized structural motifs are helix-turn- helix, zinc finger, leucine zipper, and helix-loop-helix. 5 RNA Processing
Various proteins are necessary for processing of transcribed RNAs in the nucleus. Pre-mRNA processing steps include capping at the 5' end with methylguanosine, polyadenylating the 3' end, and spUcing to remove introns. The spUceosomal complex is comprised of five small nuclear ribonucleoprotein particles (snRNPs) designated Ul, U2, U4, U5, and U6. Each snRNP contains a o single species of snRNA and about ten proteins. The RNA components of some snRNPs recognize and base-pair with intron consensus sequences. The protein components mediate sphceosome assembly and the splicing reaction. Autoantibodies to snRNP proteins are found in the blood of patients with systemic lupus erythematosus (Stryer, L. (1995) Biochemistry W.H. Freeman and Company, New York NY, p. 863). Heterogeneous nuclear ribonucleoproteins (hnRNPs) have been identified that have roles in spUcing, exporting of the mature RNAs to the cytoplasm, and mRNA translation (Biamonti, G. et al. (1998) CUn. Exp. Rheumatol. 16:317-326). Some examples of hnRNPs include the yeast proteins Hrplp, involved in cleavage and polyadenylation at the 3' end of the RNA; Cbp80p, involved in 5 capping the 5 ' end of the RNA; and Npl3p, a homolog of mammaUan hnRNP Al , involved in export of mRNA from the nucleus (Shen, E.C et al. (1998) Genes Dev. 12:679-691). HnRNPs have been shown to be important targets of the autoimmune response in rheumatic diseases (Biamonti, supra).
Many snRNP proteins, hnRNP proteins, and alternative splicing factors are characterized by an RNA recognition motif (RRM). (Reviewed in Birney, E. et al. (1993) Nucleic Acids Res. 21 :5803- 0 5816.) The RRM is about 80 amino acids in length and forms four β-strands and two α-helices arranged in an α/β sandwich. The RRM contains a core RNP-1 octapeptide motif along with surrounding conserved sequences. RNA StabiUty and Degradation
RNA heUcases alter and regulate RNA conformation and secondary structure by using energy 5 derived from ATP hydrolysis to destabiUze and unwind RNA duplexes. The most well-characterized and ubiquitous family of RNA heUcases is the DEAD-box family, so named for the conserved B-type ATP-binding motif which is diagnostic of proteins in this family. Over 40 DEAD-box heUcases have been identified in organisms as diverse as bacteria, insects, yeast, amphibians, mammals, and plants. DEAD-box heUcases function in diverse processes such as translation initiation, spUcing, ribosome o assembly, and RNA editing, transport, and stabiUty. Some DEAD-box heUcases play tissue- and stage- specific roles in spermatogenesis and embryogenesis. (Reviewed in Linder, P. et al. (1989) Nature 337:121-122.)
Overexpression of the DEAD-box 1 protein (DDX1) may play a role in the progression of neuroblastoma (Nb) and retinoblastoma (Rb) tumors. Other DEAD-box heUcases have been implicated 5 either directly or indirectly in ultraviolet Ught-induced tumors, B cell lymphoma, and myeloid maUgnancies. (Reviewed in Godbout, R. et al. (1998) J. Biol. Chem. 273:21161-21168.)
Ribonucleases (RNases) catalyze the hydrolysis of phosphodiester bonds in RNA chains, thus cleaving the RNA. For example, RNase P is a ribonucleoprotein enzyme which cleaves the 5' end of pre-tRNAs as part of their maturation process. RNase H digests the RNA strand of an RNA/DNA o hybrid. Such hybrids occur in cells invaded by retroviruses, and RNase H is an important enzyme in the retroviral repUcation cycle. RNase H domains are often found as a domain associated with reverse transcriptases. RNase activity in serum and cell extracts is elevated in a variety of cancers and infectious diseases (Schein, CH. (1997) Nat. Biotechnol. 15:529-536). Regulation of RNase activity is being investigated as a means to control tumor angiogenesis, allergic reactions, viral infection and repUcation, and fungal infections. Protein Translation
The eukaryotic ribosome is composed of a 60S (large) subunit and a 40S (small) subunit, which together form the 80S ribosome. In addition to the 18S, 28S, 5S, and 5.8S rRNAs, the ribosome 5 also contains more than fifty proteins. The ribosomal proteins have a prefix which denotes the subunit to which they belong, either L (large) or S (small). Three important sites are identified on the ribosome. The aminoacyl-tRNA site (A site) is where charged tRNAs (with the exception of the initiator-tRNA) bind on arrival at the ribosome. The peptidyl-tRNA site (P site) is where new peptide bonds are formed, as well as where the initiator tRNA binds. The exit site (E site) is where deacylated tRNAs o bind prior to their release from the ribosome. (Translation is reviewed in Stryer, L. (1995)
Biochemistry, W.H. Freeman and Company, New York NY, pp. 875-908; and Lodish, H. et al. (1995) Molecular Cell Biology, Scientific American Books, New York NY, pp. 119-138.) tRNA Charging
Protein biosynthesis depends on each amino acid forming a linkage with the appropriate tRNA. 5 The aminoacyl-tRNA synthetases are responsible for the activation and correct attachment of an amino acid with its cognate tRNA. The 20 aminoacyl-tRNA synthetase enzymes can be divided into two structural classes, Class I and Class II. Autoantibodies against aminoacyl-tRNAs are generated by patients with dermatomyositis and polymyositis, and correlate strongly with compUcating interstitial lung disease (ILD). These antibodies appear to be generated in response to viral infection, and o coxsackie virus has been used to induce experimental viral myositis in animals.
Translation Initiation
Initiation of translation can be divided into three stages. The first stage brings an initiator transfer RNA (Met-tRNAf) together with the 40S ribosomal subunit to form the 43S preinitiation complex. The second stage binds the 43S preinitiation complex to the mRNA, followed by migration of 5 the complex to the correct AUG initiation codon. The third stage brings the 60S ribosomal subunit to the 40S subunit to generate an 80S ribosome at the initiation codon. Regulation of translation primarily involves the first and second stage in the initiation process (Pain, V.M. (1996) Eur. J. Biochem. 236:747-771).
Several initiation factors, many of which contain multiple subunits, are involved in bringing an o initiator tRNA and 40S ribosomal subunit together. eIF2, a guanine nucleotide binding protein, recruits the initiator tRNA to the 40S ribosomal subunit. Only when eIF2 is bound to GTP does it associate with the initiator tRNA. eIF2B, a guanine nucleotide exchange protein, is responsible for converting eIF2 from the GDP-bound inactive form to the GTP-bound active form. Two other factors, elFl A and eIF3 bind and stabilize the 40S subunit by interacting with 18S ribosomal RNA and specific ribosomal structural proteins. eIF3 is also involved in association of the 40S ribosomal subunit with mRNA. The Met-tRNAf, elFl A, eIF3, and 40S ribosomal subunit together make up the 43 S preinitiation complex (Pain, supra).
Additional factors are required for binding of the 43 S preinitiation complex to an mRNA molecule, and the process is regulated at several levels. eIF4F is a complex consisting of three proteins: eIF4E, eIF4A, and eIF4G. eIF4E recognizes and binds to the mRNA 5 -terminal m7GTP cap, eIF4A is a bidirectional RNA-dependent heUcase, and eIF4G is a scaffolding polypeptide. eIF4G has three binding domains. The N-terminal third of eIF4G interacts with eDF4E, the central third interacts with eIF4A, and the C-terminal third interacts with eIF3 bound to the 43S preinitiation complex. Thus, eIF4G acts as a bridge between the 40S ribosomal subunit and the mRNA (Hentze, M.W. (1997) Science 275:500-501).
The abiUty of eIF4F to initiate binding of the 43S preinitiation complex is regulated by structural features of the mRNA. The mRNA molecule has an untranslated region (UTR) between the 5' cap and the AUG start codon. In some mRNAs this region forms secondary structures that impede binding of the 43S preinitiation complex. The heUcase activity of eIF4A is thought to function in removing this secondary structure to faciUtate binding of the 43S preinitiation complex (Pain, supra). Translation Elongation
Elongation is the process whereby additional amino acids are joined to the initiator methionine to form the complete polypeptide chain. The elongation factors EFlα, EFlβ γ, and EF2 are involved in elongating the polypeptide chain following initiation. EF 1 α is a GTP-binding protein. In EF 1 α' s GTP-bound form, it brings an aminoacyl-tRNA to the ribosome' s A site. The amino acid attached to the newly arrived aminoacyl-tRNA forms a peptide bond with the initiator methionine. The GTP on EFlα is hydrolyzed to GDP, and EFlα-GDP dissociates from the ribosome. EFlβ γ binds EFlα -GDP and induces the dissociation of GDP from EFlα, allowing EFlα to bind GTP and a new cycle to begin. As subsequent aminoacyl-tRNAs are brought to the ribosome, EF-G, another GTP-binding protein, catalyzes the translocation of tRNAs from the A site to the P site and finally to the E site of the ribosome. This allows the processivity of translation. Translation Termination
The release factor eRF carries out termination of translation. eRF recognizes stop codons in the mRNA, leading to the release of the polypeptide chain from the ribosome. Post-Translational Pathways
Proteins may be modified after translation by the addition of phosphate, sugar, prenyl, fatty acid, and other chemical groups. These modifications are often required for proper protein activity. Enzymes involved in post-translational modification include kinases, phosphatases, glycosyltransferases, and prenyltransferases. The conformation of proteins may also be modified after translation by the introduction and rearrangement of disulfide bonds (rearrangement catalyzed by protein disulfide isomerase), the isomerization of proline sidechains by prolyl isomerase, and by interactions with molecular chaperone proteins. Proteins may also be cleaved by proteases. Such cleavage may result in activation, inactivation, or complete degradation of the protein. Proteases include serine proteases, cysteine proteases, aspartic proteases, and metalloproteases. Signal peptidase in the endoplasmic reticulum (ER) lumen cleaves the signal peptide from membrane or secretory proteins that are imported into the ER. Ubiquitin proteases are associated with the ubiquitin conjugation system (UCS), a major pathway for the degradation of cellular proteins in eukaryotic cells and some bacteria. The UCS mediates the elimination of abnormal proteins and regulates the half-Uves of important regulatory proteins that control cellular processes such as gene transcription and cell cycle progression. In the
UCS pathway, proteins targeted for degradation are conjugated to a ubiquitin, a small heat stable protein. Proteins involved in the UCS include ubiquitin-activating enzyme, ubiquitin-conjugating enzymes, ubiquitin-ligases, and ubiquitin C-terminal hydrolases. The ubiquitinated protein is then recognized and degraded by the proteasome, a large, multisubunit proteolytic enzyme complex, and ubiquitin is released for reutilization by ubiquitin protease.
Lipid MetaboUsm
Lipids are water-insoluble, oily or greasy substances that are soluble in nonpolar solvents such as chloroform or ether. Neutral fats (triacylglycerols) serve as major fuels and energy stores. Polar
Upids, such as phosphoUpids, sphingoUpids, glycoUpids, and cholesterol, are key structural components of cell membranes.
Lipid metabolism is involved in human diseases and disorders. In the arterial disease atherosclerosis, fatty lesions form on the inside of the arterial wall. These lesions promote the loss of arterial flexibiUty and the formation of blood clots (Guyton, A.C. Textbook of Medical Physiology
(1991) W.B. Saunders Company, Philadelphia PA, pp.760-763). In Tay-Sachs disease, the GM2 ganglioside (a sphingoUpid) accumulates in lysosomes of the central nervous system due to a lack of the enzyme N-acetylhexosaminidase. Patients suffer nervous system degeneration leading to early death (Fauci, A.S. et al. (1998) Harrison's Principles of Internal Medicine McGraw-Hill, New York NY, p. 2171). The Niemann-Pick diseases are caused by defects in Upid metaboUsm. Niemann-Pick diseases types A and B are caused by accumulation of sphingomyelin (a sphingoUpid) and other Upids in the central nervous system due to a defect in the enzyme sphingomyeUnase, leading to neurodegeneration and lung disease. Niemann-Pick disease type C results from a defect in cholesterol transport, leading to the accumulation of sphingomyelin and cholesterol in lysosomes and a secondary reduction in sphingomyeUnase activity. Neurological symptoms such as grand mal seizures, ataxia, and loss of previously learned speech, manifest 1-2 years after birth. A mutation in the NPC protein, which contains a putative cholesterol-sensing domain, was found in a mouse model of Niemann-Pick disease type C (Fauci, supra, p. 2175; Loftus, S.K. et al. (1997) Science 277:232-235). (Lipid metaboUsm is reviewed in Stryer, L. (1995) Biochemistry, W.H. Freeman and Company, New York NY; Lehninger, A. (1982) Principles of Biochemistry Worth PubUshers, Inc., New York NY; and ExPASy "Biochemical Pathways" index of Boehringer Mannheim World Wide Web site.) Fatty Acid Synthesis
Fatty acids are long-chain organic acids with a single carboxyl group and a long non-polar hydrocarbon tail. Long-chain fatty acids are essential components of glycoUpids, phosphoUpids, and cholesterol, which are building blocks for biological membranes, and of triglycerides, which are biological fuel molecules. Long-chain fatty acids are also substrates for eicosanoid production, and are important in the functional modification of certain complex carbohydrates and proteins. 16-carbon and 18-carbon fatty acids are the most common. Fatty acid synthesis occurs in the cytoplasm. In the first step, acetyl-Coenzyme A (CoA) carboxylase (ACC) synthesizes malonyl-CoA from acetyl-CoA and bicarbonate. The enzymes which catalyze the remaining reactions are covalently Unked into a single polypeptide chain, referred to as the multifunctional enzyme fatty acid synthase (FAS). FAS catalyzes the synthesis of palmitate from acetyl-CoA and malonyl-CoA. FAS contains acetyl transferase, malonyl transferase, β-ketoacetyl synthase, acyl carrier protein, β-ketoacyl reductase, dehydratase, enoyl reductase, and thioesterase activities. The final product of the FAS reaction is the 16-carbon fatty acid palmitate. Further elongation, as well as unsaturation, of palmitate by accessory enzymes of theER produces the variety of long chain fatty acids required by the individual cell. These enzymes include a NADH-cytochrome b5 reductase, cytochrome b5, and a desaturase. PhosphoUpid and Triacylglvcerol Synthesis
Triacylglycerols, also known as triglycerides and neutral fats, are major energy stores in animals. Triacylglycerols are esters of glycerol with three fatty acid chains. Glyceτol-3-phosphate is produced from dihydroxyacetone phosphate by the enzyme glycerol phosphate dehydrogenase or from glycerol by glycerol kinase. Fatty acid-CoA's are produced from fatty acids by fatty acyl-CoA synthetases. Glyercol-3-phosphate is acylated with two fatty acyl-CoA's by the enzyme glycerol phosphate acyltransferase to give phosphatidate. Phosphatidate phosphatase converts phosphatidate to diacylglycerol, which is subsequently acylated to a triacylglyercol by the enzyme diglyceride acyltransferase. Phosphatidate phosphatase and diglyceride acyltransferase form a triacylglyerol synthetase complex bound to the ER membrane. A major class of phospholipids are the phosphoglycerides, which are composed of a glycerol backbone, two fatty acid chains, and a phosphorylated alcohol. Phosphoglycerides are components of cell membranes. Principal phosphoglycerides are phosphatidyl choUne, phosphatidyl ethanolamine, phosphatidyl serine, phosphatidyl inositol, and diphosphatidyl glycerol. Many enzymes involved in 5 phosphoglyceride synthesis are associated with membranes (Meyers, R.A. (1995) Molecular Biology and Biotechnology, VCH PubUshers Inc., New York NY, pp. 494-501). Phosphatidate is converted to CDP-diacylglycerol by the enzyme phosphatidate cytidylyltransferase (ExPASy ENZYME EC 2.7.7.41). Transfer of the diacylglycerol group from CDP-diacylglycerol to serine to yield phosphatidyl serine, or to inositol to yield phosphatidyl inositol, is catalyzed by the enzymes CDP-diacylglycerol- 0 serine 0-phosphatidyltransferase and CDP-diacylglycerol-inositol 3-phosphatidyltransferase, respectively (ExPASy ENZYME EC 2.7.8.8; ExPASy ENZYME EC 2.7.8.11). The enzyme phosphatidyl serine decarboxylase catalyzes the conversion of phosphatidyl serine to phosphatidyl ethanolamine, using a pyruvate cofactor (Voelker, D.R. (1997) Biochim. Biophys. Acta 1348:236-244). Phosphatidyl choUne is formed using diet-derived choUne by the reaction of CDP-choUne with 1 ,2- 5 diacylglycerol, catalyzed by diacylglycerol choUnephosphotransferase (ExPASy ENZYME 2.7.8.2). Sterol, Steroid, and Isoprenoid MetaboUsm
Cholesterol, composed of four fused hydrocarbon rings with an alcohol at one end, moderates the fluidity of membranes in which it is incoφorated. In addition, cholesterol is used in the synthesis of steroid hormones such as cortisol, progesterone, estrogen, and testosterone. Bile salts derived from o cholesterol facilitate the digestion of Upids. Cholesterol in the skin forms a barrier that prevents excess water evaporation from the body. Farnesyl and geranylgeranyl groups, which are derived from cholesterol biosynthesis intermediates, are post-translationally added to signal transduction proteins such as ras and protein-targeting proteins such as rab. These modifications are important for the activities of these proteins (Guyton, supra; Stryer, supra, pp. 279-280, 691-702, 934). 5 Mammals obtain cholesterol derived from both de novo biosynthesis and the diet. The Uver is the major site of cholesterol biosynthesis in mammals. Two acetyl-CoA molecules initially condense to form acetoacetyl-CoA, catalyzed by a thiolase. Acetoacetyl-CoA condenses with a third acetyl-CoA to form hydroxymethylglutaryl-CoA (HMG-CoA), catalyzed by HMG-CoA synthase. Conversion of HMG-CoA to cholesterol is accompUshed via a series of enzymatic steps known as the mevalonate o pathway. The rate-Umiting step is the conversion of HMG-CoA to mevalonate by HMG-CoA reductase. The drug lovastatin, a potent inhibitor of HMG-CoA reductase, is given to patients to reduce their serum cholesterol levels. Other mevalonate pathway enzymes include mevalonate kinase, phosphomevalonate kinase, diphosphomevalonate decarboxylase, isopentenyldiphosphate isomerase, dimethylallyl transferase, geranyl transferase, farnesyl-diphosphate farnesyltransferase, squalene monooxygenase, lanosterol synthase, lathosterol oxidase, and 7-dehydrocholesterol reductase. Cholesterol is used in the synthesis of steroid hormones such as cortisol, progesterone, aldosterone, estrogen, and testosterone. First, cholesterol is converted to pregnenolone by cholesterol monooxygenases. The other steroid hormones are synthesized from pregnenolone by a series of 5 enzyme-catalyzed reactions including oxidations, isomerizations, hydroxylations, reductions, and demethylations. Examples of these enzymes include steroid Δ-isomerase, 3β-hydroxy-Δ5-steroid dehydrogenase, steroid 21 -monooxygenase, steroid 19-hydroxylase, and 3β-hydroxysteroid dehydrogenase. Cholesterol is also the precursor to vitamin D.
Numerous compounds contain 5 -carbon isoprene units derived from the mevalonate pathway 0 intermediate isopentenyl pyrophosphate. Isoprenoid groups are found in vitamin K, ubiquinone, retinal, doUchol phosphate (a carrier of oUgosaccharides needed for N-linked glycosylation), and farnesyl and geranylgeranyl groups that modify proteins. Enzymes involved include farnesyl transferase, polyprenyl transferases, doUchyl phosphatase, and dolichyl kinase. SphingoUpid MetaboUsm 5 Sphingohpids are an important class of membrane lipids that contain sphingosine, a long chain amino alcohol. They are composed of one long-chain fatty acid, one polar head alcohol, and sphingosine or sphingosine derivative. The three classes of sphingolipids are sphingomyeUns, cerebrosides, and gangUosides. SphingomyeUns, which contain phosphochoUne or phosphoethanolamine as their head group, are abundant in the myeUn sheath surrounding nerve cells. o Galactocerebrosides, which contain a glucose or galactose head group, are characteristic of the brain. Other cerebrosides are found in nonneural tissues. GangUosides, whose head groups contain multiple sugar units, are abundant in the brain, but are also found in nonneural tissues.
Sphingolipids are built on a sphingosine backbone. Sphingosine is acylated to ceramide by the enzyme sphingosine acetyltransferase. Ceramide and phosphatidyl choUne are converted to 5 sphingomyeUn by the enzyme ceramide choUne phosphotransferase. Cerebrosides are synthesized by the Unkage of glucose or galactose to ceramide by a transferase. Sequential addition of sugar residues to ceramide by transferase enzymes yields gangUosides. Eicosanoid MetaboUsm
Eicosanoids, including prostaglandins, prostacycUn, thromboxanes, and leukotrienes, are 20- o carbon molecules derived from fatty acids. Eicosanoids are signaling molecules which have roles in pain, fever, and inflammation. The precursor of all eicosanoids is arachidonate, which is generated from phospholipids by phosphoUpase A2 and from diacylglycerols by diacylglycerol Upase. Leukotrienes are produced from arachidonate by the action of Upoxygenases. Prostaglandin synthase, reductases, and isomerases are responsible for the synthesis of the prostaglandins. Prostaglandins have roles in inflammation, blood flow, ion transport, synaptic transmission, and sleep. ProstacycUn and the thromboxanes are derived from a precursor prostaglandin by the action of prostacycUn synthase and thromboxane synthases, respectively. Ketone Body MetaboUsm Pairs of acetyl-CoA molecules derived from fatty acid oxidation in the liver can condense to form acetoacetyl-CoA, which subsequently forms acetoacetate, D-3-hydroxybutyrate, and acetone. These three products are known as ketone bodies. Enzymes involved in ketone body metabolism include HMG-CoA synthetase, HMG-CoA cleavage enzyme, D-3-hydroxybutyrate dehydrogenase, acetoacetate decarboxylase, and 3-ketoacyl-CoA transferase. Ketone bodies are a normal fuel supply of the heart and renal cortex. Acetoacetate produced by the liver is transported to cells where the acetoacetate is converted back to acetyl-CoA and enters the citric acid cycle. In times of starvation, ketone bodies produced from stored triacylglyerols become an important fuel source, especially for the brain. Abnormally high levels of ketone bodies are observed in diabetics. Diabetic coma can result if ketone body levels become too great. Lipid MobiUzation
Within cells, fatty acids are transported by cytoplasmic fatty acid binding proteins (Onhne MendeUan Inheritance in Man (OMIM) *134650 Fatty Acid-Binding Protein 1, Liver; FABPl). Diazepam binding inhibitor (DBI), also known as endozepine and acyl CoA-binding protein, is an endogenous γ-aminobutyric acid (GABA) receptor Ugand which is thought to down-regulate the effects of GABA. DBI binds medium- and long-chain acyl-CoA esters with very high affinity and may function as an intracellular carrier of acyl-CoA esters (OMIM *125950 Diazepam Binding Inhibitor; DBI; PROSITE PDOC00686 Acyl-CoA-binding protein signature).
Fat stored in Uver and adipose triglycerides may be released by hydrolysis and transported in the blood. Free fatty acids are transported in the blood by albumin. Triacylglycerols and cholesterol esters in the blood are transported in Upoprotein particles. The particles consist of a core of hydrophobic lipids surrounded by a shell of polar Upids and apoUpoproteins. The protein components serve in the solubilization of hydrophobic Upids and also contain cell-targeting signals. Lipoproteins include chylomicrons, chylomicron remnants, very-low-density Upoproteins (VLDL), intermediate- density Upoproteins (IDL), low-density Upoproteins (LDL), and high-density Upoproteins (HDL). There is a strong inverse correlation between the levels of plasma HDL and risk of premature coronary heart disease.
Triacylglycerols in chylomicrons and VLDL are hydrolyzed by Upoprotein Upases that line blood vessels in muscle and other tissues that use fatty acids. Cell surface LDL receptors bind LDL particles which are then internalized by endocytosis. Absence of the LDL receptor, the cause of the disease famiUal hypercholesterolemia, leads to increased plasma cholesterol levels and ultimately to atherosclerosis. Plasma cholesteryl ester transfer protein mediates the transfer of cholesteryl esters from HDL to apoUpoprotein B-containing Upoproteins. Cholesteryl ester transfer protein is important in the reverse cholesterol transport system and may play a role in atherosclerosis (Yamashita, S. et al. (1997) Curr. Opin. Lipidol. 8:101-110). Macrophage scavenger receptors, which bind and internaUze modified Upoproteins, play a role in lipid transport and may contribute to atherosclerosis (Greaves, D.R. et al. (1998) Curr. Opin. Lipidol. 9:425-432).
Proteins involved in cholesterol uptake and biosynthesis are tightly regulated in response to cellular cholesterol levels. The sterol regulatory element binding protein (SREBP) is a sterol-responsive transcription factor. Under normal cholesterol conditions, SREBP resides in the ER membrane. When cholesterol levels are low, a regulated cleavage of SREBP occurs which releases the extracellular domain of the protein. This cleaved domain is then transported to the nucleus where it activates the transcription of the LDL receptor gene, and genes encoding enzymes of cholesterol synthesis, by binding the sterol regulatory element (SRE) upstream of the genes (Yang, J. et al. (1995) J. Biol. Chem. 270:12152-12161). Regulation of cholesterol uptake and biosynthesis also occurs via the oxysterol- binding protein (OSBP). OSBP is a high-affinity intracellular receptor for a variety of oxysterols that down-regulate cholesterol synthesis and stimulate cholesterol esterification (Lagace, T.A. et al. (1997) Biochem. J. 326:205-213). Beta-oxidation Mitochondrial and peroxisomal beta-oxidation enzymes degrade saturated and unsaturated fatty acids by sequential removal of two-carbon units from CoA-activated fatty acids. The main beta- oxidation pathway degrades both saturated and unsaturated fatty acids while the auxiUary pathway performs additional steps required for the degradation of unsaturated fatty acids.
The pathways of mitochondrial and peroxisomal beta-oxidation use similar enzymes, but have different substrate specificities and functions. Mitochondria oxidize short-, medium-, and long-chain fatty acids to produce energy for cells. Mitochondrial beta-oxidation is a major energy source for cardiac and skeletal muscle. In liver, it provides ketone bodies to the peripheral circulation when glucose levels are low as in starvation, endurance exercise, and diabetes (Eaton, S. et al. (1996) Biochem. J. 320:345-357). Peroxisomes oxidize medium-, long-, and very-long-chain fatty acids, dicarboxyUc fatty acids, branched fatty acids, prostaglandins, xenobiotics, and bile acid intermediates. The chief roles of peroxisomal beta-oxidation are to shorten toxic UpophiUc carboxylic acids to faciUtate their excretion and to shorten very-long-chain fatty acids prior to mitochondrial beta-oxidation (Mannaerts, G.P. and P.P. van Veldhoven ( 1993) Biochimie 75:147-158).
Enzymes involved in beta-oxidation include acyl CoA synthetase, carnitine acyltransferase, acyl CoA dehydrogenases, enoyl CoA hydratases, L-3-hydroxyacyl CoA dehydrogenase, β-ketothiolase, 2,4-dienoyl CoA reductase, and isomerase. Lipid Cleavage and Degradation
Triglycerides are hydrolyzed to fatty acids and glycerol by Upases. LysophosphoUpases 5 (LPLs) are widely distributed enzymes that metaboUze intracellular Upids, and occur in numerous isoforms. Small isoforms, approximately 15-30 kD, function as hydrolases; large isoforms, those exceeding 60 kD, function both as hydrolases and transacylases. A particular substrate for LPLs, lysophosphatidylchoUne, causes lysis of cell membranes when it is formed or imported into a cell. LPLs are regulated by Upid factors including acylcarnitine, arachidonic acid, and phosphatidic acid. l o These Upid factors are signaling molecules important in numerous pathways, including the inflammatory response. (Anderson, R. et al. (1994) Toxicol. Appl. Pharmacol. 125:176-183; Selle, H. et al. (1993); Eur. J. Biochem. 212:411-416.)
The secretory phosphoUpase A2 (PLA2) superfamily comprises a number of heterogeneous enzymes whose common feature is to hydrolyze the sn-2 fatty acid acyl ester bond of
15 phosphoglycerides. Hydrolysis of the glycerophosphoUpids releases free fatty acids and lysophosphoUpids. PLA2 activity generates precursors for the biosynthesis of biologically active Upids, hydroxy fatty acids, and platelet-activating factor. PLA2 hydrolysis of the sn-2 ester bond in phosphoUpids generates free fatty acids, such as arachidonic acid and lysophosphoUpids. Carbon and Carbohydrate MetaboUsm
20 Carbohydrates, including sugars or saccharides, starch, and cellulose, are aldehyde or ketone compounds with multiple hydroxyl groups. The importance of carbohydrate metaboUsm is demonstrated by the sensitive regulatory system in place for maintenance of blood glucose levels. Two pancreatic hormones, insulin and glucagon, promote increased glucose uptake and storage by cells, and increased glucose release from cells, respectively. Carbohydrates have three important roles in
25 mammaUan cells. First, carbohydrates are used as energy stores, fuels, and metabolic intermediates. Carbohydrates are broken down to form energy in glycolysis and are stored as glycogen for later use. Second, the sugars deoxyribose and ribose form part of the structural support of DNA and RNA, respectively. Third, carbohydrate modifications are added to secreted and membrane proteins and Upids as they traverse the secretory pathway. Cell surface carbohydrate-containing macromolecules,
3 o including glycoproteins, glycoUpids, and transmembrane proteoglycans, mediate adhesion with other cells and with components of the extracellular matrix. The extracellular matrix is comprised of diverse glycoproteins, glycosaminoglycans (GAGs), and carbohydrate-binding proteins which are secreted from the cell and assembled into an organized meshwork in close association with the cell surface. The interaction of the cell with the surrounding matrix profoundly influences cell shape, strength, flexibility, motiUty, and adhesion. These dynamic properties are intimately associated with signal transduction pathways contrαlUng cell proUferation and differentiation, tissue construction, and embryonic development.
Carbohydrate metaboUsm is altered in several disorders including diabetes melUtus, 5 hyperglycemia, hypoglycemia, galactosemia, galactokinase deficiency, and UDP-galactose-4-epimerase deficiency (Fauci, AS. et al. (1998) Harrison's Principles of Internal Medicine. McGraw-Hill, New York NY, pp. 2208-2209). Altered carbohydrate metaboUsm is associated with cancer. Reduced GAG and proteoglycan expression is associated with human lung carcinomas (Nackaerts, K. et al. (1997) Int. J. Cancer 74:335-345). The carbohydrate determinants sialyl Lewis A and sialyl Lewis X are 0 frequently expressed on human cancer cells (Kannagi, R. (1997) Glycoconj. J. 14:577-584).
Alterations of the N-Unked carbohydrate core structure of cell surface glycoproteins are Unked to colon and pancreatic cancers (Schwarz, R.E. et al. (1996) Cancer Lett. 107:285-291). Reduced expression of the Sda blood group carbohydrate structure in cell surface glycoUpids and glycoproteins is observed in gastrointestinal cancer (Dobi, T. et al. (1996) Int. J. Cancer 67:626-663). (Carbon and 5 carbohydrate metaboUsm is reviewed in Stryer, L. (1995) Biochemistry W.H. Freeman and Company, New York NY; Lehninger, A.L. (1982) Principles of Biochemistry Worth PubUshers Inc., New York NY; and Lodish, H. et al. (1995) Molecular Cell Biology Scientific American Books, New York NY.) Glycolysis
Enzymes of the glycolytic pathway convert the sugar glucose to pyruvate while simultaneously o producing ATP. The pathway also provides building blocks for the synthesis of cellular components such as long-chain fatty acids. After glycolysis, pyrvuate is converted to acetyl-Coenzyme A, which, in aerobic organisms, enters the citric acid cycle. Glycolytic enzymes include hexokinase, phosphoglucose isomerase, phosphofructokinase, aldolase, triose phosphate isomerase, glyceraldehyde 3-phosphate dehydrogenase, phosphoglycerate kinase, phosphoglyceromutase, enolase, and pyruvate kinase. Of 5 these, phosphofructokinase, hexokinase, and pyruvate kinase are important in regulating the rate of glycolysis. Gluconeogenesis
Gluconeogenesis is the synthesis of glucose from noncarbohydrate precursors such as lactate and amino acids. The pathway, which functions mainly in times of starvation and intense exercise, o occurs mostly in the liver and kidney. Responsible enzymes include pyruvate carboxylase, phosphoenolpyruvate carboxykinase, fructose 1,6-bisphosphatase, and glucose-6-phosphatase. Pentose Phosphate Pathway
Pentose phosphate pathway enzymes are responsible for generating the reducing agent NADPH, while at the same time oxidizing glucose-6-phosphate to ribose-5 -phosphate. Ribose-5- phosphate and its derivatives become part of important biological molecules such as ATP, Coenzyme A, NAD+, FAD, RNA, and DNA. The pentose phosphate pathway has both oxidative and non- oxidative branches. The oxidative branch steps, which are catalyzed by the enzymes glucose-6- phosphate dehydrogenase, lactonase, and 6-phosphogluconate dehydrogenase, convert glucose-6- 5 phosphate and NADP+ to ribulose-6-phosphate and NADPH. The non-oxidative branch steps, which are catalyzed by the enzymes phosphopentose isomerase, phosphopentose epimerase, transketolase, and transaldolase, allow the interconversion of three-, four-, five-, six-, and seven-carbon sugars. Glucouronate MetaboUsm
Glucuronate is a monosaccharide which, in the form of D-glucuronic acid, is found in the o GAGs chondroitin and dermatan. D-glucuronic acid is also important in the detoxification and excretion of foreign organic compounds such as phenol. Enzymes involved in glucuronate metaboUsm include UDP-glucose dehydrogenase and glucuronate reductase. Disaccharide MetaboUsm
Disaccharides must be hydrolyzed to monosaccharides to be digested. Lactose, a disaccharide 5 found in milk, is hydrolyzed to galactose and glucose by the enzyme lactase. Maltose is derived from plant starch and is hydrolyzed to glucose by the enzyme maltase. Sucrose is derived from plants and is hydrolyzed to glucose and fructose by the enzyme sucrase. Trehalose, a disaccharide found mainly in insects and mushrooms, is hydrolyzed to glucose by the enzyme trehalase (OMIM *275360 Trehalase; Ruf, J. et al. (1990) J. Biol. Chem. 265:15034-15039). Lactase, maltase, sucrase, and trehalase are o bound to mucosal cells lining the small intestine, where they participate in the digestion of dietary disaccharides. The enzyme lactose synthetase, composed of the catalytic subunit galactosyltransferase and the modifier subunit α-lactalbumin, converts UDP-galactose and glucose to lactose in the mammary glands.
Glvcogen, Starch, and Chitin Metabolism 5 Glycogen is the storage form of carbohydrates in mammals. MobiUzation of glycogen maintains glucose levels between meals and during muscular activity. Glycogen is stored mainly in the Uver and in skeletal muscle in the form of cytoplasmic granules. These granules contain enzymes that catalyze the synthesis and degradation of glycogen, as well as enzymes that regulate these processes. Enzymes that catalyze the degradation of glycogen include glycogen phosphorylase, a transferase, α- 0 1 ,6-glucosidase, and phosphoglucomutase. Enzymes that catalyze the synthesis of glycogen include UDP-glucose pyrophosphorylase, glycogen synthetase, a branching enzyme, and nucleoside diphosphokinase. The enzymes of glycogen synthesis and degradation are tightly regulated by the hormones insuUn, glucagon, and epinephrine. Starch, a plant-derived polysaccharide, is hydrolyzed to maltose, maltotriose, and α-dextrin by α-amylase, an enzyme secreted by the saUvary glands and pancreas. Chitin is a polysaccharide found in insects and Crustacea. A chitotriosidase is secreted by macrophages and may play a role in the degradation of chitin-containing pathogens (Boot, R.G. et al. (1995) J. Biol. Chem. 270:26252-26256). Peptidoglvcans and Grvcosaminoglvcans 5 Glycosaminoglycans (GAGs) are anionic Unear unbranched polysaccharides composed of repetitive disaccharide units. These repetitive units contain a derivative of an amino sugar, either glucosamine or galactosamine. GAGs exist free or as part of proteoglycans, large molecules composed of a core protein attached to one or more GAGs. GAGs are found on the cell surface, inside cells, and in the extracellular matrix. Changes in GAG levels are associated with several autoimmune diseases o including autoimmune thyroid disease, autoimmune diabetes melUtus, and systemic lupus erythematosus (Hansen, C et al. (1996) CUn. Exp. Rheum. 14 (Suppl. 15):S59-S67). GAGs include chondroitin sulfate, keratan sulfate, heparin, heparan sulfate, dermatan sulfate, and hyaluronan.
The GAG hyaluronan (HA) is found in the extracellular matrix of many cells, especially in soft connective tissues, and is abundant in synovial fluid (PitsilUdes, AA. et al. (1993) Int. J. Exp. Pathol. 5 74:27-34). HA seems to play important roles in cell regulation, development, and differentiation (Laurent, T.C. and J.R. Fraser (1992) FASEB J. 6:2397-2404). Hyaluronidase is an enzyme that degrades HA to oUgosaccharides. Hyaluronidases may function in cell adhesion, infection, angiogenesis, signal transduction, reproduction, cancer, and inflammation.
Proteoglycans, also known as peptidoglycans, are found in the extracellular matrix of o connective tissues such as cartilage and are essential for distributing the load in weight-bearing joints.
Cell-surface-attached proteoglycans anchor cells to the extracellular matrix. Both extracellular and cell-surface proteoglycans bind growth factors, facilitating their binding to cell-surface receptors and subsequent triggering of signal transduction pathways. Amino Acid and Nitrogen Metabolism 5 NH4 + is assimilated into amino acids by the actions of two enzymes, glutamate dehydrogenase and glutamine synthetase. The carbon skeletons of amino acids come from the intermediates of glycolysis, the pentose phosphate pathway, or the citric acid cycle. Of the twenty amino acids used in proteins, humans can synthesize only thirteen (nonessential amino acids). The remaining nine must come from the diet (essential amino acids). Enzymes involved in nonessential o amino acid biosynthesis include glutamate kinase dehydrogenase, pyrroline carboxylate reductase, asparagine synthetase, phenylalanine oxygenase, methionine adenosyltransferase, adenosylhomocysteinase, cystathionine β-synthase, cystathionine γ-lyase, phosphoglycerate dehydrogenase, phosphoserine transaminase, phosphoserine phosphatase, serine hydroxylmethyltransferase, and glycine synthase. Metabolism of amino acids takes place almost entirely in the liver, where the amino group is removed by aminotransferases (transaminases), for example, alanine aminotransferase. The amino group is transferred to α-ketoglutarate to form glutamate. Glutamate dehydrogenase converts glutamate to NH4 + and α-ketoglutarate. NH4 +is converted to urea by the urea cycle which is 5 catalyzed by the enzymes arginase, ornithine transcarbamoylase, arginosuccinate synthetase, and arginosuccinase. Carbamoyl phosphate synthetase is also involved in urea formation. Enzymes involved in the metabolism of the carbon skeleton of amino acids include serine dehydratase, asparaginase, glutaminase, propionyl CoA carboxylase, methylmalonyl CoA mutase, branched-chain α-keto dehydrogenase complex, isovaleryl CoA dehydrogenase, β-methylcrotonyl CoA carboxylase, o phenylalanine hydroxylase, p-hydroxylphenylpyruvate hydroxylase, and homogentisate oxidase.
Polyamines, which include spermidine, putrescine, and spermine, bind tightly to nucleic acids and are abundant in rapidly proliferating cells. Enzymes involved in polyamine synthesis include ornithine decarboxylase.
Diseases involved in amino acid and nitrogen metabolism include hyperammonemia, 5 carbamoyl phosphate synthetase deficiency, urea cycle enzyme deficiencies, methylmalonic aciduria, maple syrup disease, alcaptonuria, and phenylketonuria. Energy MetaboUsm
Cells derive energy from metabolism of ingested compounds that may be roughly categorized as carbohydrates, fats, or proteins. Energy is also stored in polymers such as triglycerides (fats) and o glycogen (carbohydrates). MetaboUsm proceeds along separate reaction pathways connected by key intermediates such as acetyl coenzyme A (acetyl-CoA). MetaboUc pathways feature anaerobic and aerobic degradation, coupled with the energy-requiring reactions such as phosphorylation of adenosine diphosphate (ADP) to the triphosphate (ATP) or analogous phosphorylations of guanosine (GDP/GTP), uridine (UDP/UTP), or cytidine (CDP/CTP). Subsequent dephosphorylation of the 5 triphosphate drives reactions needed for cell maintenance, growth, and proliferation.
Digestive enzymes convert carbohydrates and sugars to glucose; fructose and galactose are converted in the liver to glucose. Enzymes involved in these conversions include galactose- 1- phosphate uridyl transferase and UDP-galactose-4 epimerase. In the cytoplasm, glycolysis converts glucose to pyruvate in a series of reactions coupled to ATP synthesis. o Pyruvate is transported into the mitochondria and converted to acetyl-CoA for oxidation via the citric acid cycle, involving pyruvate dehydrogenase components, dihydrolipoyl transacetylase, and dihydrolipoyl dehydrogenase. Enzymes involved in the citric acid cycle include: citrate synthetase, aconitases, isocitrate dehydrogenase, alpha-ketoglutarate dehydrogenase complex including transsuccinylases, succinyl CoA synthetase, succinate dehydrogenase, fumarases, and malate 5 dehydrogenase. Acetyl CoA is oxidized to C02 with concomitant formation of NADH, FADH2, and GTP. In oxidative phosphorylation, the transport of electrons from NADH and FADH2 to oxygen by dehydrogenases is coupled to the synthesis of ATP from ADP and P; by the F(Ψl ATPase complex in the mitochondrial inner membrane. Enzyme complexes responsible for electron transport and ATP synthesis include the F^ ATPase complex, ubiquinone(CoQ)-cytochrome c reductase, ubiquinone 5 reductase, cytochrome b, cytochrome Cj, FeS protein, and cytochrome c oxidase.
Triglycerides are hydrolyzed to fatty acids and glycerol by Upases. Glycerol is then phosphorylated to glycerol-3-phosphate by glycerol kinase and glycerol phosphate dehydrogenase, and degraded by the glycolysis. Fatty acids are transported into the mitochondria as fatty acyl- carnitine esters and undergo oxidative degradation. 0 In addition to metaboUc disorders such as diabetes and obesity, disorders of energy metabolism are associated with cancers (Dorward, A. et al. (1997) J. Bioenerg. Biomembr. 29:385- 392), autism (Lombard, J. (1998) Med. Hypotheses 50:497-500), neurodegenerative disorders (Alexi, T. et al. (1998) Neuroreport 9:R57-64), and neuromuscular disorders (DiMauro, S. et al. (1998) Biochim. Biophys. Acta 1366:199-210). The myocardium is heavily dependent on oxidative 5 metabolism, so metabolic dysfunction often leads to heart disease (DiMauro, S. and M. Hirano (1998) Curr. Opin. Cardiol. 13:190-197).
For a review of energy metaboUsm enzymes and intermediates, see Stryer, L. et al. (1995) Biochemistry. W.H. Freeman and Co., San Francisco CA, pp. 443-652. For a review of energy metabolism regulation, seeLodish, H. et al. (1995) Molecular Cell Biology, Scientific American o Books, New York NY, pp. 744-770. Cofactor Metabolism
Cofactors, including coenzymes and prosthetic groups, are small molecular weight inorganic or organic compounds that are required for the action of an enzyme. Many cofactors contain vitamins as a component. Cofactors include thiamine pyrophosphate, flavin adenine dinucleotide, flavin 5 mononucleotide, nicotinamide adenine dinucleotide, pyridoxal phosphate, coenzyme A, tetrahydrofolate, Upoamide, and heme. The vitamins biotin and cobalamin are associated with enzymes as well. Heme, a prosthetic group found in myoglobin and hemoglobin, consists of protopoφhyrin group bound to iron. Poφhyrin groups contain four substituted pyrroles covalently joined in a ring, often with a bound metal atom. Enzymes involved in poφhyrin synthesis include δ- o aminolevulinate synthase, δ-aminolevulinate dehydrase, poφhobilinogen deaminase, and cosynthase.
Deficiencies in heme formation cause poφhyrias. Heme is broken down as a part of erythrocyte turnover. Enzymes involved in heme degradation include heme oxygenase and biliverdin reductase. Iron is a required cofactor for many enzymes. Besides the heme-containing enzymes, iron is found in iron-sulfur clusters in proteins including aconitase, succinate dehydrogenase, and NADH-Q 5 reductase. Iron is transported in the blood by the protein transferrin. Binding of transferrin to the transferrin receptor on cell surfaces allows uptake by receptor mediated endocytosis. Cytosolic iron is bound to ferritin protein.
A molybdenum-containing cofactor (molybdopterin) is found in enzymes including sulfite oxidase, xanthine dehydrogenase, and aldehyde oxidase. Molybdopterin biosynthesis is performed by 5 two molybdenum cofactor synthesizing enzymes. Deficiencies in these enzymes cause mental retardation and lens dislocation. Other diseases caused by defects in cofactor metabolism include pernicious anemia and methylmalonic aciduria. Secretion and Trafficking
Eukaryotic cells are bound by a Upid bilayer membrane and subdivided into functionally o distinct, membrane bound compartments. The membranes maintain the essential differences between the cytosol, the extracellular environment, and the lumenal space of each intracellular organelle. As Upid membranes are highly impermeable to most polar molecules, transport of essential nutrients, metaboUc waste products, cell signaUng molecules, macromolecules and proteins across lipid membranes and between organelles must be mediated by a variety of transport-associated molecules. 5 Protein Trafficking
In eukaryotes, some proteins are synthesized on ER-bound ribosomes, co-translationally imported into the ER, deUvered from the ER to the Golgi complex for post-translational processing and sorting, and transported from the Golgi to specific intracellular and extracellular destinations. All cells possess a constitutive transport process which maintains homeostasis between the cell and its o environment. In many differentiated cell types, the basic machinery is modified to carry out specific transport functions. For example, in endocrine glands, hormones and other secreted proteins are packaged into secretory granules for regulated exocytosis to the cell exterior. In macrophage, foreign extracellular material is engulfed (phagocytosis) and deUvered to lysosomes for degradation. In fat and muscle cells, glucose transporters are stored in vesicles which fuse with the plasma membrane only in 5 response to insulin stimulation.
The Secretory Pathway
Synthesis of most integral membrane proteins, secreted proteins, and proteins destined for the lumen of a particular organelle occurs on ER-bound ribosomes. These proteins are co-translationally imported into the ER. The proteins leave the ER via membrane-bound vesicles which bud off the ER at o specific sites and fuse with each other (homotypic fusion) to form the ER-Golgi Intermediate
Compartment (ERGIC). The ERGIC matures progressively through the cis, medial, and trans cisternal stacks of the Golgi, modifying the enzyme composition by retrograde transport of specific Golgi enzymes. In this way, proteins moving through the Golgi undergo post-translational modification, such as glycosylation. The final Golgi compartment is the Trans-Golgi Network (TGN), where both membrane and lumenal proteins are sorted for their final destination. Transport vesicles destined for intracellular compartments, such as the lysosome, bud off the TGN. What remains is a secretory vesicle which contains proteins destined for the plasma membrane, such as receptors, adhesion molecules, and ion channels, and secretory proteins, such as hormones, neurotransmitters, and digestive 5 enzymes. Secretory vesicles eventually fuse with the plasma membrane (GUck, B.S. and V. Malhotra (1998) Cell 95:883-889).
The secretory process can be constitutive or regulated. Most cells have a constitutive pathway for secretion, whereby vesicles derived from maturation of the TGN require no specific signal to fuse with the plasma membrane. In many cells, such as endocrine cells, digestive cells, and neurons, vesicle o pools derived from the TGN collect in the cytoplasm and do not fuse with the plasma membrane until they are directed to by a specific signal.
Endocytosis
Endocytosis, wherein cells internaUze material from the extracellular environment, is essential for transmission of neuronal, metabolic, and proUferative signals; uptake of many essential nutrients; 5 and defense against invading organisms. Most cells exhibit two forms of endocytosis. The first, phagocytosis, is an actin-driven process exempUfied in macrophage and neutrophils. Material to be endocytosed contacts numerous cell surface receptors which stimulate the plasma membrane to extend and surround the particle, enclosing it in a membrane-bound phagosome. In the mammalian immune system, IgG-coated particles bind Fc receptors on the surface of phagocytic leukocytes. Activation of 0 the Fc receptors initiates a signal cascade involving src-family cytosoUc kinases and the monomeric
GTP-binding (G) protein Rho. The resulting actin reorganization leads to phagocytosis of the particle.
This process is an important component of the humoral immune response, allowing the processing and presentation of bacterial-derived peptides to antigen-specific T-lymphocytes.
The second form of endocytosis, pinocytosis, is a more generaUzed uptake of material from the 5 external miUeu. Like phagocytosis, pinocytosis is activated by Ugand binding to cell surface receptors.
Activation of individual receptors stimulates an internal response that includes coalescence of the receptor-Ugand complexes and formation of clathrin-coated pits. Imagination of the plasma membrane at clathrin-coated pits produces an endocytic vesicle within the cell cytoplasm. These vesicles undergo homotypic fusion to form an early endosomal (EE) compartment. The tubulovesicular EE serves as a o sorting site for incoming material. ATP-driven proton pumps in the EE membrane lowers the pH of the
EE lumen (pH 6.3-6.8). The acidic environment causes many Ugands to dissociate from their receptors. The receptors, along with membrane and other integral membrane proteins, are recycled back to the plasma membrane by budding off the tubular extensions of the EE in recycUng vesicles (RV). This selective removal of recycled components produces a carrier vesicle containing Ugand and other material from the external environment. The carrier vesicle fuses with TGN-derived vesicles which contain hydrolytic enzymes. The acidic environment of the resulting late endosome (LE) activates the hydrolytic enzymes which degrade the Ugands and other material. As digestion takes place, the LE fuses with the lysosome where digestion is completed (MeUman, I. (1996) Annu. Rev. Cell Dev. Biol. 5 12:575-625).
RecycUng vesicles may return directly to the plasma membrane. Receptors internaUzed and returned directly to the plasma membrane have a turnover rate of 2-3 minutes. Some RVs undergo microtubule-directed relocation to a perinuclear site, from which they then return to the plasma membrane. Receptors following this route have a turnover rate of 5-10 minutes. Still other RVs are 0 retained within the cell until an appropriate signal is received (Mellman, supra; and James, D.E. et al. (1994) Trends Cell Biol. 4:120-126). Vesicle Formation
Several steps in the transit of material along the secretory and endocytic pathways require the formation of transport vesicles. Specifically, vesicles form at the transitional endoplasmic reticulum 5 (tER), the rim of Golgi cisternae, the face of the Trans-Golgi Network (TGN), the plasma membrane (PM), and tubular extensions of the endosomes. The process begins with the budding of a vesicle out of the donor membrane. The membrane-bound vesicle contains proteins to be transported and is surrounded by a protective coat made up of protein subunits recruited from the cytosol. The initial budding and coating processes are controlled by a cytosotic ras-Uke GTP-binding protein, ADP- o ribosylating factor (Arf), and adapter proteins (AP). Different isoforms of both Arf and AP are involved at different sites of budding. Another small G-protein, dynamin, forms a ring complex around the neck of the forming vesicle and may provide the mechanochemical force to accompUsh the final step of the budding process. The coated vesicle complex is then transported through the cytosol. During the transport process, Arf-bound GTP is hydrolyzed to GDP and the coat dissociates from the transport 5 vesicle (West, M.A. et al. (1997) J. Cell Biol. 138:1239-1254). Two different classes of coat protein have also been identified. Clathrin coats form on the TGN and PM surfaces, whereas coatomer or COP coats form on the ER and Golgi. COP coats can further be distinguished as COPI, involved in retrograde traffic through the Golgi and from the Golgi to the ER, and COPII, involved in anterograde traffic from the ER to the Golgi (Mellman, supra). The COP coat consists of two major components, a o G-protein (Arf or Sar) and coat protomer (coatomer). Coatomer is an equimolar complex of seven proteins, termed alpha-, beta-, beta'-, gamma-, delta-, epsilon- and zeta-COP. (Harter, C. and F.T. Wieland (1998) Proc. Natl. Acad. Sci. USA 95:11649-11654.) Membrane Fusion
Transport vesicles undergo homotypic or heterotypic fusion in the secretory and endocytotic pathways. Molecules required for appropriate targeting and fusion of vesicles with their target membrane include proteins incoφorated in the vesicle membrane, the target membrane, and proteins recruited from the cytosol. During budding of the vesicle from the donor compartment, an integral membrane protein, VAMP (vesicle-associated membrane protein) is incoφorated into the vesicle. Soon 5 after the vesicle uncoats, a cytosoUc prenylated GTP-binding protein, Rab (a member of the Ras superfamily), is inserted into the vesicle membrane. GTP-bound Rab proteins are directed into nascent transport vesicles where they interact with VAMP. Following vesicle transport, GTPase activating proteins (GAPs) in the target membrane convert Rab proteins to the GDP-bound form. A cytosoUc protein, guanine-nucleotide dissociation inhibitor (GDI) helps return GDP-bound Rab proteins to their o membrane of origin. Several Rab isoforms have been identified and appear to associate with specific compartments within the cell. Rab proteins appear to play a role in mediating the function of a viral gene, Rev, which is essential for repUcation of HIV-1, the virus responsible for AIDS (Flavell, R.A. et al. (1996) Proc. Natl. Acad. Sci. USA 93:4421-4424).
Docking of the transport vesicle with the target membrane involves the formation of a complex 5 between the vesicle SNAP receptor (v-SNARE), target membrane (t-) SNAREs, and certain other membrane and cytosoUc proteins. Many of these other proteins have been identified although their exact functions in the docking complex remain uncertain (Tellam, J.T. et al. (1995) J. Biol. Chem. 270:5857-63; and Hata, Y and T.C Sudhof (1995) J. Biol. Chem. 270:13022-28). N-ethylmaleimide sensitive factor (NSF) and soluble NSF-attacbment protein (α-SNAP and β-SNAP) are two such o proteins that are conserved from yeast to man and function in most intracellular membrane fusion reactions. Seel represents a family of yeast proteins that function at many different stages in the secretory pathway including membrane fusion. Recently, mammaUan homologs of Seel, called Munc-18 proteins, have been identified (Katagiri, H. et al. (1995) J. Biol. Chem. 270:4963-4966; Hata et al. supra). 5 The SNARE complex involves three SNARE molecules, one in the vesicular membrane and two in the target membrane. Synaptotagmin is an integral membrane protein in the synaptic vesicle which associates with the t-SNARE syntaxin in the docking complex. Synaptotagmin binds calcium in a complex with negatively charged phosphoUpids, which allows the cytosoUc SNAP protein to displace synaptotagmin from syntaxin and fusion to occur. Thus, synaptotagmin is a negative regulator of o fusion in the neuron (Littleton, J.T. et al. (1993) Cell 74:1125-1134). The most abundant membrane protein of synaptic vesicles appears to be the glycoprotein synaptophysin, a 38 kDa protein with four transmembrane domains.
Specificity between a vesicle and its target is derived from the v-SNARE, t-SNAREs, and associated proteins involved. Different isoforms of SNAREs and Rabs show distinct cellular and subcellular distributions. VAMP-1/synaptobrevin, membrane-anchored synaptosome-associated protein of 25 kDa (SNAP- 25), syntaxin-1 , Rab3A, Rabl5, and Rab23 are predominantly expressed in the brain and nervous system. Different syntaxin, VAMP, and Rab proteins are associated with distinct subcellular compartments and their vesicular carriers. 5 Nuclear Transport
Transport of proteins and RNA between the nucleus and the cytoplasm occurs through nuclear pore complexes (NPCs). NPC-mediated transport occurs in both directions through the nuclear envelope. All nuclear proteins are imported from the cytoplasm, their site of synthesis. tRNA and mRNA are exported from the nucleus, their site of synthesis, to the cytoplasm, their site of function. o Processing of small nuclear RNAs involves export into the cytoplasm, assembly with proteins and modifications such as hypermethylation to produce small nuclear ribonuclear proteins (snRNPs), and subsequent import of the snRNPs back into the nucleus. The assembly of ribosomes requires the initial import of ribosomal proteins from the cytoplasm, their incoφoration with RNA into ribosomal subunits, and export back to the cytoplasm. (Gorlich, D. and I.W. Mattaj (1996) Science 271:1513- 5 1518.)
The transport of proteins and mRNAs across the NPC is selective, dependent on nuclear localization signals, and generally requires association with nuclear transport factors. Nuclear locaUzation signals (NLS) consist of short stretches of amino acids enriched in basic residues. NLS are found on proteins that are targeted to the nucleus, such as the glucocorticoid receptor. The NLS is o recognized by the NLS receptor, importin, which then interacts with the monomeric GTP-binding protein Raa This NLS protein/receptor/Ran complex navigates the nuclear pore with the help of the homodimeric protein nuclear transport factor 2 (NTF2). NTF2 binds the GDP-bound form of Ran and to multiple proteins of the nuclear pore complex containing FXFG repeat motifs, such as p62. (Paschal, B. et al. (1997) J. Biol. Chem. 272:21534-21539; and Wong, D.H. et al. (1997) Mol. Cell 5 Biol. 17:3755-3767). Some proteins are dissociated before nuclear mRNAs are transported across the
NPC while others are dissociated shortly after nuclear mRNA transport across the NPC and are reimported into the nucleus. Disease Correlation
The etiology of numerous human diseases and disorders can be attributed to defects in the o transport or secretion of proteins. For example, abnormal hormonal secretion is Unked to disorders such as diabetes insipidus (vasopressin), hyper- and hypoglycemia (insuUn, glucagon), Grave's disease and goiter (thyroid hormone), and Cushing's and Addison's diseases (adrenocorticotropic hormone, ACTH). Moreover, cancer cells secrete excessive amounts of hormones or other biologically active peptides. Disorders related to excessive secretion of biologically active peptides by tumor cells include fasting hypoglycemia due to increased insuUn secretion from insuUnoma-islet cell tumors; hypertension due to increased epinephrine and norepinephrine secreted from pheochromocytomas of the adrenal medulla and sympathetic paragangUa; and carcinoid syndrome, which is characterized by abdominal cramps, diarrhea, and valvular heart disease caused by excessive amounts of vasoactive substances such as serotonin, bradykinin, histamine, prostaglandins, and polypeptide hormones, secreted from intestinal tumors. Biologically active peptides that are ectopically synthesized in and secreted from tumor cells include ACTH and vasopressin (lung and pancreatic cancers); parathyroid hormone (lung and bladder cancers); calcitonin (lung and breast cancers); and thyroid-stimulating hormone (medullary thyroid carcinoma). Such peptides may be useful as diagnostic markers for tumorigenesis (Schwartz, M.Z. (1997) Semin. Pediatr. Surg. 3:141-146; and Said, S.I. and G.R. Faloona (1975) N. Engl. J. Med. 293:155-160).
Defective nuclear transport may play a role in cancer. The BRCAl protein contains three potential NLSs which interact with importin alpha, and is transported into the nucleus by the importin/NPC pathway. In breast cancer cells the BRCAl protein is aberrantly locahzed in the cytoplasm. The mislocation of the BRCAl protein in breast cancer cells may be due to a defect in the NPC nuclear import pathway (Chen, CF. et al. (1996) J. Biol. Chem. 271:32863-32868).
It has been suggested that in some breast cancers, the tumor-suppressing activity of p53 is inactivated by the sequestration of the protein in the cytoplasm, away from its site of action in the cell nucleus. Cytoplasmic wild-type p53 was also found in human cervical carcinoma cell Unes. (Moll, U.M. et al. (1992) Proc. Natl. Acad. Sci. USA 89:7262-7266; and Liang, X.H. et al. (1993) Oncogene 8:2645-2652.) Environmental Responses
Organisms respond to the environment by a number of pathways. Heat shock proteins, including hsp 70, hsp60, hsp90, and hsp 40, assist organisms in coping with heat damage to cellular proteins.
Aquaporins (AQP) are channels that transport water and, in some cases, nonionic small solutes such as urea and glycerol. Water movement is important for a number of physiological processes including renal fluid filtration, aqueous humor generation in the eye, cerebrospinal fluid production in the brain, and appropriate hydration of the lung. Aquaporins are members of the major intrinsic protein (MIP) family of membrane transporters (King, L.S. and P. Agre (1996) Annu. Rev. Physiol. 58:619- 648; Ishibashi, K. et al. (1997) J. Biol. Chem. 272:20782-20786). The study of aquaporins may have relevance to understanding edema formation and fluid balance in both normal physiology and disease states (King, supra). Mutations in AQP2 cause autosomal recessive nephrogenic diabetes insipidus (OMIM *107777 Aquaporin 2; AQP2). Reduced AQP4 expression in skeletal muscle may be associated with Duchenne muscular dystrophy (Frigeri, A et al. (1998) J. CUn. Invest. 102:695-703). Mutations in AQPO cause autosomal dominant cataracts in the mouse (OMIM *154050 Major Intrinsic Protein of Lens Fiber; MIP).
The metallothioneins (MTs) are a group of small (61 amino acids), cysteine-rich proteins that bind heavy metals such as cadmium, zinc, mercury, lead, and copper and are thought to play a role in metal detoxification or the metaboUsm and homeostasis of metals. Arsenite-resistance proteins have been identified in hamsters that are resistant to toxic levels of arsenite (Rossman, T.G. et al. (1997) Mutat. Res. 386:307-314).
Humans respond to light and odors by specific protein pathways. Proteins involved in light perception include rhodopsin, fransducin, and cGMP phosphodiesterase. Proteins involved in odor perception include multiple olfactory receptors. Other proteins are important in human Circadian rhythms and responses to wounds. Immunity and Host Defense
Al vertebrates have developed sophisticated and complex immune systems that provide protection from viral, bacterial, fungal and parasitic infections. Included in these systems are the processes of humoral immunity, the complement cascade and the inflammatory response (Paul, W.E. (1993) Fundamental Immunology. Raven Press, Ltd., New York NY, pp.1-20).
The cellular components of the humoral immune system include six different types of leukocytes: monocytes, lymphocytes, polymoφhonuclear granulocytes (consisting of neutrophils, eosinophils, and basophils) and plasma cells. Additionally, fragments of megakaryocytes, a seventh type of white blood cell in the bone marrow, occur in large numbers in the blood as platelets.
Leukocytes are formed from two stem cell lineages in bone marrow. The myeloid stem cell line produces granulocytes and monocytes and, the lymphoid stem cell produces lymphocytes. Lymphoid cells travel to the thymus, spleen and lymph nodes, where they mature and differentiate into lymphocytes. Leukocytes are responsible for defending the body against invading pathogens. Neutrophils and monocytes attack invading bacteria, viruses, and other pathogens and destroy them by phagocytosis. Monocytes enter tissues and differentiate into macrophages which are extremely phagocytic. Lymphocytes and plasma cells are a part of the immune system which recognizes specific foreign molecules and organisms and inactivates them, as well as signals other cells to attack the invaders.
Granulocytes and monocytes are formed and stored in the bone marrow until needed. Megakaryocytes are produced in bone marrow, where they fragment into platelets and are released into the bloodstream. The main function of platelets is to activate the blood clotting mechanism. Lymphocytes and plasma cells are produced in various lymphogenous organs, including the lymph nodes, spleen, thymus, and tonsils.
Both neutrophils and macrophages exhibit chemotaxis towards sites of inflammation. Tissue inflammation in response to pathogen invasion results in production of chemo-attractants for leukocytes, such as endotoxins or other bacterial products, prostaglandins, and products of leukocytes 5 or platelets. .
Basophils participate in the release of the chemicals involved in the inflammatory process. The main function of basophils is secretion of these chemicals to such a degree that they have been referred to as "unicellular endocrine glands". A distinct aspect of basophilic secretion is that the contents of granules go directly into the extracellular environment, not into vacuoles as occurs with 0 neutrophils, eosinophils and monocytes. Basophils have receptors for the Fc fragment of immunoglobulin E (IgE) that are not present on other leukocytes. CrossUnking of membrane IgE with anti-IgE or other ligands triggers degranulation.
Eosinophils are bi- or multi-nucleated white blood cells which contain eosinophiUc granules. Their plasma membrane is characterized by Ig receptors, particularly IgG and IgE. Generally, 5 eosinophils are stored in the bone marrow until recruited for use at a site of inflammation or invasion. They have specific functions in parasitic infections and allergic reactions, and are thought to detoxify some of the substances released by mast cells and basophils which cause inflammation. Additionally, they phagocytize antigen-antibody complexes and further help prevent spread of the inflammation.
Macrophages are monocytes that have left the blood stream to settle in tissue. Once o monocytes have migrated into tissues, they do not re-enter the bloodstream. The mononuclear phagocyte system is comprised of precursor cells in the bone marrow, monocytes in circulation, and macrophages in tissues. The system is capable of very fast and extensive phagocytosis. A macrophage may phagocytize over 100 bacteria, digest them and extrude residues, and then survive for many more months. Macrophages are also capable of ingesting large particles, including red 5 blood cells and malarial parasites. They increase several-fold in size and transform into macrophages that are characteristic of the tissue they have entered, surviving in tissues for several months.
Mononuclear phagocytes are essential in defending the body against invasion by foreign pathogens, particularly intracellular microorganisms such as M. tuberculosis, listeria, leishmania and toxoplasma. Macrophages can also control the growth of tumorous cells, via both phagocytosis and o secretion of hydrolytic enzymes. Another important function of macrophages is that of processing antigen and presenting them in a biochemically modified form to lymphocytes.
The immune system responds to invading microorganisms in two major ways: antibody production and cell mediated responses. Antibodies are immunoglobulin proteins produced by B-lymphocytes which bind to specific antigens and cause inactivation or promote destruction of the 5 antigen by other cells. Cell -mediated immune responses involve T-lymphocytes (T cells) that react with foreign antigen on the surface of infected host cells. Depending on the type of T cell, the infected cell is either killed or signals are secreted which activate macrophages and other cells to destroy the infected cell (Paul, supra).
T-lymphocytes originate in the bone marrow or liver in fetuses. Precursor cells migrate via 5 the blood to the thymus, where they are processed to mature into T-lymphocytes. This processing is crucial because of positive and negative selection of T cells that will react with foreign antigen and not with self molecules. After processing, T cells continuously circulate in the blood and secondary lymphoid tissues, such as lymph nodes, spleen, certain epithelium-associated tissues in the gastrointestinal tract, respiratory tract and skin. When T-lymphocytes are presented with the o complementary antigen, they are stimulated to proliferate and release large numbers of activated T cells into the lymph system and the blood system. These activated T cells can survive and circulate for several days. At the same time, T memory cells are created, which remain in the lymphoid tissue for months or years. Upon subsequent exposure to that specific antigen, these memory cells will respond more rapidly and with a stronger response than induced by the original antigen. This creates 5 an "immunological memory" that can provide immunity for years.
There are two major types of T cells: cytotoxic T cells destroy infected host cells, and helper T cells activate other white blood cells via chemical signals. One class of helper cell, TH1 , activates macrophages to destroy ingested microorganisms, while another, TH2, stimulates the production of antibodies by B cells. o Cytotoxic T cells directly attack the infected target cell. In virus-infected cells, peptides derived from viral proteins are generated by the proteasome. These peptides are transported into the ER by the transporter associated with antigen processing (TAP) (Pamer, E. and P. Cresswell (1998) Annu. Rev. Immunol. 16:323-358). Once inside the ER, the peptides bind MHC I chains, and the peptide/MHC I complex is transported to the cell surface. Receptors on the surface of T cells bind to 5 antigen presented on cell surface MHC molecules. Once activated by binding to antigen, T cells secrete γ-interferon, a signal molecule that induces the expression of genes necessary for presenting viral (or other) antigens to cytotoxic T cells. Cytotoxic T cells kill the infected cell by stimulating programmed cell death.
Helper T cells constitute up to 75% of the total T cell population. They regulate the immune o functions by producing a variety of lymphokines that act on other cells in the immune system and on bone marrow. Among these lymphokines are: interleukins-2,3,4,5,6; granulocyte-monocyte colony stimulating factor, and γ-interferon.
Helper T cells are required for most B cells to respond to antigen. When an activated helper cell contacts a B cell, its centrosome and Golgi apparatus become oriented toward the B cell, aiding 5 the directing of signal molecules, such as transmembrane-bound protein called CD40 ligand, onto the B cell surface to interact with the CD40 transmembrane protein. Secreted signals also help B cells to proliferate and mature and, in some cases, to switch the class of antibody being produced.
B-lymphocytes (B cells) produce antibodies which react with specific antigenic proteins presented by pathogens. Once activated, B cells become filled with extensive rough endoplasmic 5 reticulum and are known as plasma cells. As with T cells, interaction of B cells with antigen stimulates proliferation of only those B cells which produce antibody specific to that antigen. There are five classes of antibodies, known as immunoglobulins, which together comprise about 20% of total plasma protein. Each class mediates a characteristic biological response after antigen binding. Upon activation by specific antigen B cells switch from making membrane-bound antibody to o secretion of that antibody.
Antibodies, or immunoglobulins (Ig), are the founding members of the Ig superfamily and the central components of the humoral immune response. Antibodies are either expressed on the surface of B cells or secreted by B cells into the circulation. Antibodies bind and neutralize blood-borne foreign antigens. The prototypical antibody is a tetramer consisting of two identical heavy 5 polypeptide chains (H-chains) and two identical light polypeptide chains (L-chains) interlinked by disulfide bonds. This arrangement confers the characteristic Y-shape to antibody molecules. Antibodies are classified based on their H-chain composition. The five antibody classes, IgA, IgD, IgE, IgG and IgM, are defined by the a, δ, e, γ, and μ H-chain types. There are two types of L- chains, K and λ, either of which may associate as a pair with any H-chain pair. IgG, the most o common class of antibody found in the circulation, is tetrameric, while the other classes of antibodies are generally variants or multimers of this basic structure.
H-chains and L-chains each contain an N-terminal variable region and a C-terminal constant region. Both H-chains and L-chains contain repeated Ig domains. For example, a typical H-chain contains four Ig domains, three of which occur within the constant region and one of which occurs 5 within the variable region and contributes to the formation of the antigen recognition site. Likewise, a typical L-chain contains two Ig domains, one of which occurs within the constant region and one of which occurs within the variable region. In addition, H chains such as μ have been shown to associate with other polypeptides during differentiation of the B cell.
Antibodies can be described in terms of their two main functional domains. Antigen o recognition is mediated by the Fab (antigen binding fragment) region of the antibody, while effector functions are mediated by the Fc (crystallizable fragment) region. Binding of antibody to an antigen, such as a bacterium, triggers the destruction of the antigen by phagocytic white blood cells such as macrophages and neutrophils. These cells express surface receptors that specifically bind to the antibody Fc region and allow the phagocytic cells to engulf, ingest, and degrade the antibody-bound 5 antigen. The Fc receptors expressed by phagocytic cells are single-pass transmembrane glycoproteins of about 300 to 400 amino acids (Sears, D.W. et al. (1990) J. Immunol. 144:371-378). The extracellular portion of the Fc receptor typically contains two or three Ig domains.
Diseases which cause over- or under-abundance of any one type of leukocyte usually result in the entire immune defense system becoming involved. A well-known autoimmune disease is AIDS (Acquired Immunodeficiency Syndrome) where the number of helper T cells is depleted, leaving the patient susceptible to infection by microorganisms and parasites. Another widespread medical condition attributable to the immune system is that of allergic reactions to certain antigens. Allergic reactions include: hay fever, asthma, anaphylaxis, and urticaria (hives). Leukemias are an excess production of white blood cells, to the point where a major portion of the body's metaboUc resources are directed solely at proUferation of white blood cells, leaving other tissues to starve. Leukopenia or agranulocytosis occurs when the bone marrow stops producing white blood cells. This leaves the body unprotected against foreign microorganisms, including those which normally inhabit skin, mucous membranes, and gastrointestinal tract. If all white blood cell production stops completely, infection will occur within two days and death may follow only 1 to 4 days later. Impaired phagocytosis occurs in several diseases, including monocytic leukemia, systemic lupus, and granulomatous disease. In such a situation, macrophages can phagocytize normally, but the enveloped organism is not killed. A defect in the plasma membrane enzyme which converts oxygen to lethally reactive forms results in abscess formation in liver, lungs, spleen, lymph nodes, and beneath the skin. Eosinophilia is an excess of eosinophils commonly observed in patients with allergies (hay fever, asthma), allergic reactions to drugs, rheumatoid arthritis, and cancers (Hodgkin's disease, lung, and liver cancer) (Isselbacher, KJ. et al. (1994) Harrison's Principles of Internal Medicine, McGraw-Hill, Inc., New York NY).
Host defense is further augmented by the complement system. The complement system serves as an effector system and is involved in infectious agent recognition. It can function as an independent immune network or in conjunction with other humoral immune responses. The complement system is comprised of numerous plasma and membrane proteins that act in a cascade of reaction sequences whereby one component activates the next. The result is a rapid and amplified response to infection through either an inflammatory response or increased phagocytosis.
The complement system has more than 30 protein components which can be divided into functional groupings including modified serine proteases, membrane-binding proteins and regulators of complement activation. Activation occurs through two different pathways the classical and the alternative. Both pathways serve to destroy infectious agents through distinct triggering mechanisms that eventually merge with the involvement of the component C3.
The classical pathway requires antibody binding to infectious agent antigens. The antibodies serve to define the target and initiate the complement system cascade, culminating in the destruction of the infectious agent. In this pathway, since the antibody guides initiation of the process, the complement can be seen as an effector arm of the humoral immune system.
The alternative pathway of the complement system does not require the presence of preexisting antibodies for targeting infectious agent destruction. Rather, this pathway, through low levels of an activated component, remains constantly primed and provides surveillance in the non- immune host to enable targeting and destruction of infectious agents. In this case foreign material triggers the cascade, thereby facilitating phagocytosis or lysis (Paul, supra, pp.918-919).
Another important component of host defense is the process of inflammation. Inflammatory responses are divided into four categories on the basis of pathology and include allergic inflammation, cytotoxic antibody mediated inflammation, immune complex mediated inflammation and monocyte mediated inflammation. Inflammation manifests as a combination of each of these forms with one predominating.
Alergic acute inflammation is observed in individuals wherein specific antigens stimulate IgE antibody production. Mast cells and basophils are subsequently activated by the attachment of antigen- IgE complexes, resulting in the release of cytoplasmic granule contents such as histamine. The products of activated mast cells can increase vascular permeability and constrict the smooth muscle of breathing passages, resulting in anaphylaxis or asthma. Acute inflammation is also mediated by cytotoxic antibodies and can result in the destruction of tissue through the binding of complement-fixing antibodies to cells. The responsible antibodies are of the IgG or IgM types. Resultant clinical disorders include autoimmune hemolytic anemia and thrombocytopenia as associated with systemic lupus erythematosis.
Immune complex mediated acute inflammation involves the IgG or IgM antibody types which combine with antigen to activate the complement cascade. When such immune complexes bind to neutrophils and macrophages they activate the respiratory burst to form protein- and vessel- damaging agents such as hydrogen peroxide, hydroxyl radical, hypochlorous acid, and chloramines.
Clinical manifestations include rheumatoid arthritis and systemic lupus erythematosus.
In chronic inflammation or delayed-type hypersensitivity, macrophages are activated and process antigen for presentation to T cells that subsequently produce lymphokines and monokines. This type of inflammatory response is likely important for defense against intracellular parasites and certain viruses. Clinical associations include, granulomatous disease, tuberculosis, leprosy, and sarcoidosis (Paul, W.E., supra, pp.1017-1018).
Extracellular Information Transmission Molecules
SEQ ID NO:9 encodes, for example, an extracellular information transmission molecule. Intercellular communication is essential for the growth and survival of multicellular organisms, and in particular, for the function of the endocrine, nervous, and immune systems. In addition, intercellular communication is critical for developmental processes such as tissue construction and organogenesis, in which cell proliferation, cell differentiation, and moφhogenesis must be spatially and temporally regulated in a precise and coordinated manner. Cells communicate 5 with one another through the secretion and uptake of diverse types of signaling molecules such as hormones, growth factors, neuropeptides, and cytokines. Hormones
Hormones are signaUng molecules that coordinately regulate basic physiological processes from embryogenesis throughout adulthood. These processes include metaboUsm, respiration, o reproduction, excretion, fetal tissue differentiation and organogenesis, growth and development, homeostasis, and the stress response. Hormonal secretions and the nervous system are tightly integrated and interdependent. Hormones are secreted by endocrine glands, primarily the hypothalamus and pituitary, the thyroid and parathyroid, the pancreas, the adrenal glands, and the ovaries and testes. The secretion of hormones into the circulation is tightly controlled. Hormones are often 5 secreted in diurnal, pulsatile, and cyclic patterns. Hormone secretion is regulated by perturbations in blood biochemistry, by other upstream-acting hormones, by neural impulses, and by negative feedback loops. Blood hormone concentrations are constantly monitored and adjusted to maintain optimal, steady-state levels. Once secreted, hormones act only on those target cells that express specific receptors. o Most disorders of the endocrine system are caused by either hyposecretion or hypersecretion of hormones. Hyposecretion often occurs when a hormone's gland of origin is damaged or otherwise impaired. Hypersecretion often results from the proliferation of tumors derived from hormone-secreting cells. Inappropriate hormone levels may also be caused by defects in regulatory feedback loops or in the processing of hormone precursors. Endocrine malfunction may also occur when the target cell fails 5 to respond to the hormone.
Hormones can be classified biochemically as polypeptides, steroids, eicosanoids, or amines. Polypeptides, which include diverse hormones such as insuUn and growth hormone, vary in size and function and are often synthesized as inactive precursors that are processed intracellularly into mature, active forms. Amines, which include epinephrine and dopamine, are amino acid derivatives that o function in neuroendocrine signaUng. Steroids, which include the cholesterol-derived hormones estrogen and testosterone, function in sexual development and reproduction. Eicosanoids, which include prostaglandins and prostacycUns, are fatty acid derivatives that function in a variety of processes. Most polypeptides and some amines are soluble in the circulation where they are highly susceptible to proteolytic degradation within seconds after their secretion. Steroids and Upids are insoluble and must be transported in the circulation by carrier proteins. The following discussion will focus primarily on polypeptide hormones.
Hormones secreted by the hypothalamus and pituitary gland play a critical role in endocrine function by coordinately regulating hormonal secretions from other endocrine glands in response to neural signals. Hypothalamic hormones include thyrotropin-releasing hormone, gonadotropin-releasing hormone, somatostatin, growth-hormone releasing factor, corticotropin-releasing hormone, substance P, dopamine, and prolactin-releasing hormone. These hormones directly regulate the secretion of hormones from the anterior lobe of the pituitary. Hormones secreted by the anterior pituitary include adrenocorticotropic hormone (ACTH), melanocyte-stimulating hormone, somatotropic hormones such as growth hormone and prolactin, glycoprotein hormones such as thyroid-stimulating hormone, luteinizing hormone (LH), and folUcle-stimulating hormone (FSH), β-Upotropin, and β-endoφhins. These hormones regulate hormonal secretions from the thyroid, pancreas, and adrenal glands, and act directly on the reproductive organs to stimulate ovulation and spermatogenesis. The posterior pituitary synthesizes and secretes antidiuretic hormone (ADH, vasopressin) and oxytocin. Disorders of the hypothalamus and pituitary often result from lesions such as primary brain tumors, adenomas, infarction associated with pregnancy, hypophysectomy, aneurysms, vascular malformations, thrombosis, infections, immunological disorders, and compUcations due to head trauma. Such disorders have profound effects on the function of other endocrine glands. Disorders associated with hypopituitarism include hypogonadism, Sheehan syndrome, diabetes insipidus, Kallman's disease, Hand-Schuller-Christian disease, Letterer-Siwe disease, sarcoidosis, empty sella syndrome, and dwarfism. Disorders associated with hypeφituitarism include acromegaly, giantism, and syndrome of inappropriate ADH secretion (SIADH), often caused by benign adenomas.
Hormones secreted by the thyroid and parathyroid primarily control metabolic rates and the regulation of serum calcium levels, respectively. Thyroid hormones include calcitonin, somatostatin, and thyroid hormone. The parathyroid secretes parathyroid hormone. Disorders associated with hypothyroidism include goiter, myxedema, acute thyroiditis associated with bacterial infection, subacute thyroiditis associated with viral infection, autoimmune thyroiditis (Hashimoto's disease), and cretinism. Disorders associated with hyperthyroidism include thyrotoxicosis and its various forms, Grave's disease, pretibial myxedema, toxic multinodular goiter, thyroid carcinoma, and Plummer's disease. Disorders associated with hypeφarathyroidism include Conn disease (chronic hypercalemia) leading to bone resoφtion and parathyroid hypeφlasia.
Hormones secreted by the pancreas regulate blood glucose levels by modulating the rates of carbohydrate, fat, and protein metaboUsm. Pancreatic hormones include insuUn, glucagon, amyUn, γ- aminobutyric acid, gastrin, somatostatin, and pancreatic polypeptide. The principal disorder associated with pancreatic dysfunction is diabetes melUtus caused by insufficient insuUn activity. Diabetes melUtus is generally classified as either Type I (insulin-dependent, juvenile diabetes) or Type II (non- insuUn-dependent, adult diabetes). The treatment of both forms by insuUn replacement therapy is well known. Diabetes melUtus often leads to acute complications such as hypoglycemia (insulin shock), 5 coma, diabetic ketoacidosis, lactic acidosis, and chronic compUcations leading to disorders of the eye, kidney, skin, bone, joint, cardiovascular system, nervous system, and to decreased resistance to infection.
The anatomy, physiology, and diseases related to hormonal function are reviewed in McCance, K.L. and S.E. Huether (1994) Pathophysiology: The Biological Basis for Disease in Adults and 0 Children, Mosby-Year Book, Inc., St. Louis MO; Greenspan, F.S. and J.D. Baxter (1994) Basic and CUnical Endocrinology, Appleton and Lange, East Norwalk CT. Growth Factors
Growth factors are secreted proteins that mediate intercellular communication. UnUke hormones, which travel great distances via the circulatory system, most growth factors are primarily 5 local mediators that act on neighboring cells. Most growth factors contain a hydrophobic N-terminal signal peptide sequence which directs the growth factor into the secretory pathway. Most growth factors also undergo post-translational modifications within the secretory pathway. These modifications can include proteolysis, glycosylation, phosphorylation, and intramolecular disulfide bond formation. Once secreted, growth factors bind to specific receptors on the surfaces of neighboring o target cells, and the bound receptors trigger intracellular signal transduction pathways. These signal transduction pathways eUcit specific cellular responses in the target cells. These responses can include the modulation of gene expression and the stimulation or inhibition of cell division, cell differentiation, and cell motility.
Growth factors fall into at least two broad and overlapping classes. The broadest class 5 includes the large polypeptide growth factors, which are wide-ranging in their effects. These factors include epidermal growth factor (EGF), fibroblast growth factor (FGF), transforming growth factor-β (TGF-β), insulin-like growth factor (IGF), nerve growth factor (NGF), and platelet-derived growth factor (PDGF), each defining a family of numerous related factors. The large polypeptide growth factors, with the exception of NGF, act as mitogens on diverse cell types to stimulate wound healing, o bone synthesis and remodeling, extracellular matrix synthesis, and proliferation of epithelial, epidermal, and connective tissues. Members of the TGF-β, EGF, and FGF famines also function as inductive signals in the differentiation of embryonic tissue. NGF functions specifically as a neurotrophic factor, promoting neuronal growth and differentiation.
Another class of growth factors includes the hematopoietic growth factors, which are narrow in their target specificity. These factors stimulate the proUferation and differentiation of blood cells such as B-lymphocytes, T-lymphocytes, erythrocytes, platelets, eosinophils, basophils, neutrophils, macrophages, and their stem cell precursors. These factors include the colony-stimulating factors (G- CSF, M-CSF, GM-CSF, and CSF1-3), erythropoietin, and the cytokines. The cytokines are speciaUzed hematopoietic factors secreted by cells of the immune system and are discussed in detail below.
Growth factors play critical roles in neoplastic transformation of cells in vitro and in tumor progression in vivo. Overexpression of the large polypeptide growth factors promotes the proliferation and transformation of cells in culture. Inappropriate expression of these growth factors by tumor cells in vivo may contribute to tumor vascularization and metastasis. Inappropriate activity of hematopoietic growth factors can result in anemias, leukemias, and lymphomas. Moreover, growth factors are both structurally and functionally related to oncoproteins, the potentially cancer-causing products of proto-oncogenes. Certain FGF and PDGF family members are themselves homologous to oncoproteins, whereas receptors for some members of the EGF, NGF, and FGF families are encoded by proto-oncogenes. Growth factors also affect the transcriptional regulation of both proto-oncogenes and oncosuppressor genes (Pimentel, E. (1994) Handbook of Growth Factors, CRC Press, Ann Arbor MI; McKay, I. and I. Leigh, eds. (1993) Growth Factors: A Practical Approach, Oxford University Press, New York NY; Habenicht, A., ed. (1990) Growth Factors. Differentiation Factors, and Cytokines. Springer- Verlag, New York NY).
In addition, some of the large polypeptide growth factors play crucial roles in the induction of the primordial germ layers in the developing embryo. This induction ultimately results in the formation of the embryonic mesoderm, ectoderm, and endoderm which in turn provide the framework for the entire adult body plan. Disruption of this inductive process would be catastrophic to embryonic development. Small Peptide Factors - Neuropeptides and Vasomediators Neuropeptides and vasomediators (NP/VM) comprise a family of small peptide factors, typically of 20 amino acids or less. These factors generally function in neuronal excitation and inhibition of vasoconstriction vasodilation, muscle contraction, and hormonal secretions from the brain and other endocrine tissues. Included in this family are neuropeptides and neuropeptide hormones such as bombesin, neuropeptide Y, neurotensin, neuromedin N, melanocortins, opioids, galanin, somatostatin, tachykinins, urotensin II and related peptides involved in smooth muscle stimulation, vasopressin, vasoactive intestinal peptide, and circulatory system-borne signaUng molecules such as angiotensin, complement, calcitonin, endotheUns, formyl-methionyl peptides, glucagon, cholecystokinin, gastrin, and many of the peptide hormones discussed above. NP/VMs can transduce signals directly, modulate the activity or release of other neurotransmitters and hormones, and act as catalytic enzymes in signaUng cascades. The effects of NP/VMs range from extremely brief to long-lasting. (Reviewed in Martin, CR. et al. (1985) Endocrine Physiology, Oxford University Press, New York NY, pp. 57-62.) Cytokines 5 Cytokines comprise a family of signaling molecules that modulate the immune system and the inflammatory response. Cytokines are usually secreted by leukocytes, or white blood cells, in response to injury or infection. Cytokines function as growth and differentiation factors that act primarily on cells of the immune system such as B- and T-lymphocytes, monocytes, macrophages, and granulocytes. Like other signaUng molecules, cytokines bind to specific plasma membrane receptors and trigger o intracellular signal transduction pathways which alter gene expression patterns. There is considerable potential for the use of cytokines in the treatment of inflammation and immune system disorders.
Cytokine structure and function have been extensively characterized in vitro. Most cytokines are small polypeptides of about 30 kilodaltons or less. Over 50 cytokines have been identified from human and rodent sources. Examples of cytokine subfamiUes include the interferons (IFN-α, -β, and - 5 γ), the interleukins (ILl-ILl 3), the tumor necrosis factors (TNF-α and -β), and the chemokines. Many cytokines have been produced using recombinant DNA techniques, and the activities of individual cytokines have been determined in vitro. These activities include regulation of leukocyte proliferation, differentiation, and motiUty.
The activity of an individual cytokine in vitro may not reflect the full scope of that cytokine' s o activity in vivo. Cytokines are not expressed individually in vivo but are instead expressed in combination with a multitude of other cytokines when the organism is challenged with a stimulus. Together, these cytokines collectively modulate the immune response in a manner appropriate for that particular stimulus. Therefore, the physiological activity of a cytokine is determined by the stimulus itself and by complex interactive networks among co-expressed cytokines which may demonstrate both 5 synergistic and antagonistic relationships.
Chemokines comprise a cytokine subfamily with over 30 members. (Reviewed in Wells, T. N.C. and M.C Peitsch (1997) J. Leukoc. Biol. 61:545-550.) Chemokines were initially identified as chemotactic proteins that recruit monocytes and macrophages to sites of inflammation. Recent evidence indicates that chemokines may also play key roles in hematopoiesis and HIV-1 infection. Chemokines 0 are small proteins which range from about 6-15 kilodaltons in molecular weight. Chemokines are further classified as C, CC, CXC, or CX3C based on the number and position of critical cysteine residues. The CC chemokines, for example, each contain a conserved motif consisting of two consecutive cysteines followed by two additional cysteines which occur downstream at 24- and 16- residue intervals, respectively (ExPASy PROSITE database, documents PS00472 and PDOC00434). The presence and spacing of these four cysteine residues are highly conserved, whereas the intervening residues diverge significantly. However, a conserved tyrosine located about 15 residues downstream of the cysteine doublet seems to be important for chemotactic activity. Most of the human genes encoding CC chemokines are clustered on chromosome 17, although there are a few examples of CC chemokine genes that map elsewhere. Other chemokines include lymphotactin (C chemokine); macrophage chemotactic and activating factor (MCAF/MCP-1; CC chemokine); platelet factor 4 and IL-8 (CXC chemokines); and fractalkine and neurofractin (CX3C chemokines). (Reviewed in Luster, A.D. (1998) N. Engl. J. Med. 338:436-445.)
Receptor Molecules
SEQ ID NO:10 and SEQ ID NO:l 1 encode, for example, receptor molecules. The term receptor describes proteins that specifically recognize other molecules. The category is broad and includes proteins with a variety of functions. The bulk of receptors are cell surface proteins which bind extracellular ligands and produce cellular responses in the areas of growth, differentiation, endocytosis, and immune response. Other receptors faciUtate the selective transport of proteins out of the endoplasmic reticulum and locaUze enzymes to particular locations in the cell. The term may also be apphed to proteins which act as receptors for Ugands with known or unknown chemical composition and which interact with other cellular components. For example, the steroid hormone receptors bind to and regulate transcription of DNA. Regulation of cell proUferation, differentiation, and migration is important for the formation and function of tissues. Regulatory proteins such as growth factors coordinately control these cellular processes and act as mediators in cell-cell signaUng pathways. Growth factors are secreted proteins that bind to specific cell-surface receptors on target cells. The bound receptors trigger intracellular signal transduction pathways which activate various downstream effectors that regulate gene expression, cell division, cell differentiation, cell motility, and other cellular processes.
Cell surface receptors are typically integral plasma membrane proteins. These receptors recognize hormones such as catecholamines; peptide hormones; growth and differentiation factors; small peptide factors such as thyrotropin-releasing hormone; galanin, somatostatin, and tachykinins; and circulatory system-borne signaUng molecules. Cell surface receptors on immune system cells recognize antigens, antibodies, and major histocompatibiUty complex (MHC)-bound peptides. Other cell surface receptors bind ligands to be internaUzed by the cell. This receptor-mediated endocytosis functions in the uptake of low density Upoproteins (LDL), transferrin, glucose- or mannose-terminal glycoproteins, galactose-terminal glycoproteins, immunoglobulins, phosphovitellogenins, fibrin, proteinase-inhibitor complexes, plasminogen activators, and thrombospondin (Lodish, H. et al. (1995) Molecular Cell Biology, Scientific American Books, New York NY, p. 723; Mikhailenko, I. et al. (1997) J. Biol. Chem. 272:6784-6791). Receptor Protein Kinases
Many growth factor receptors, including receptors for epidermal growth factor, 5 platelet-derived growth factor, fibroblast growth factor, as well as the growth modulator α-thrombin, contain intrinsic protein kinase activities. When growth factor binds to the receptor, it triggers the autophosphorylation of a serine, threonine, or tyrosine residue on the receptor. These phosphorylated sites are recognition sites for the binding of other cytoplasmic signaUng proteins. These proteins participate in signaUng pathways that eventually link the initial receptor activation at the cell surface to 0 the activation of a specific intracellular target molecule. In the case of tyrosine residue autophosphorylation, these signaUng proteins contain a common domain referred to as a Src homology (SH) domain. SH2 domains and SH3 domains are found in phosphoUpase C-γ, PI-3-K p85 regulatory subunit, Ras-GTPase activating protein, and ppόO 0 (Lowenstein, E.J. et al. (1992) Cell 70:431-442). The cytokine family of receptors share a different common binding domain and include transmembrane 5 receptors for growth hormone (GH), interleukins, erythropoietin, and prolactia
Other receptors and second messenger-binding proteins have intrinsic serine/threonine protein kinase activity. These include activin/TGF-β/BMP-superfamily receptors, calcium- and diacylglycerol- activated/phosphoUpid-dependant protein kinase (PK-C), and RNA-dependant protein kinase (PK-R). In addition, other serine/threonine protein kinases, including nematode Twitchin, have fibronectin-Uke, o immunoglobuUn C2-Uke domains.
G-Protein Coupled Receptors
G-protein coupled receptors (GPCRs) are integral membrane proteins characterized by the presence of seven hydrophobic transmembrane domains which span the plasma membrane and form a bundle of antiparallel alpha (α) heUces. These proteins range in size from under 400 to over 1000 5 amino acids (Strosberg, AD. (1991) Eur. J. Biochem. 196:1-10; CoughUn, S.R. (1994) Curr. Opin. Cell Biol. 6:191-197). The amino-terminus of the GPCR is extracellular, of variable length and often glycosylated; the carboxy-terminus is cytoplasmic and generally phosphorylated. Extracellular loops of the GPCR alternate with intracellular loops and Unk the transmembrane domains. The most conserved domains of GPCRs are the transmembrane domains and the first two cytoplasmic loops. The o transmembrane domains account for structural and functional features of the receptor. In most cases, the bundle of α heUces forms a binding pocket. In addition, the extracellular N-terminal segment or one or more of the three extracellular loops may also participate in Ugand binding. Ligand binding activates the receptor by inducing a conformational change in intracellular portions of the receptor. The activated receptor, in turn, interacts with an intracellular heterotrimeric guanine nucleotide binding (G) protein complex which mediates further intracellular signaling activities, generally the production of second messengers such as cycUc AMP (cAMP), phosphoUpase C, inositol triphosphate, or interactions with ion channel proteins (Baldwin, J.M. (1994) Curr. Opin. Cell Biol. 6:180-190).
GPCRs include those for acetylchoUne, adenosine, epinephrine and norepinephrine, bombesin, 5 bradykinin, chemokines, dopamine, endotheUn, γ-aminobutyric acid (GABA), folUcle-stimulating hormone (FSH), glutamate, gonadotropin-releasing hormone (GnRH), hepatocyte growth factor, histamine, leukotrienes, melanocortins, neuropeptide Y, opioid peptides, opsins, prostanoids, serotonin, somatostatin, tachykinins, thrombin, thyrotropin-releasing hormone (TRH), vasoactive intestinal polypeptide family, vasopressin and oxytocin, and oφhan receptors. o GPCR mutations, which may cause loss of function or constitutive activation, have been associated with numerous human diseases (CoughUn, supra). For instance, retinitis pigmentosa may arise from mutations in the rhodopsin gene. Rhodopsin is the retinal photoreceptor which is located within the discs of the eye rod cell. Parma, J. et al. (1993, Nature 365:649-651) report that somatic activating mutations in the thyrotropin receptor cause hyperfunctioning thyroid adenomas and suggest 5 that certain GPCRs susceptible to constitutive activation may behave as protooncogenes. Nuclear Receptors
Nuclear receptors bind small molecules such as hormones or second messengers, leading to increased receptor-binding affinity to specific chromosomal DNA elements. In addition the affinity for other nuclear proteins may also be altered. Such binding and protein-protein interactions may regulate o and modulate gene expression. Examples of such receptors include the steroid hormone receptors family, the retinoic acid receptors family, and the thyroid hormone receptors family. Ligand-Gated Receptor Ion Channels
Ligand-gated receptor ion channels fall into two categories. The first category, extracellular Ugand-gated receptor ion channels (ELGs), rapidly transduce neurotransmitter-binding events into 5 electrical signals, such as fast synaptic neurotransmission. ELG function is regulated by post- translational modification. The second category, intracellular ligand-gated receptor ion channels (ILGs), are activated by many intracellular second messengers and do not require post-translational modifications) to effect a channel-opening response.
ELGs depolarize excitable cells to the threshold of action potential generation. In non-excitable o cells, ELGs permit a Umited calcium ion-influx during the presence of agonist. ELGs include channels directly gated by neurotransmitters such as acetylchoUne, L-glutamate, glycine, ATP, serotonin, GABA, and histamine. ELG genes encode proteins having strong structural and functional similarities. ILGs are encoded by distinct and unrelated gene famiUes and include receptors for cAMP, cGMP, calcium ions, ATP, and metaboUtes of arachidonic acid. Macrophage Scavenger Receptors
Macrophage scavenger receptors with broad ligand specificity may participate in the binding of low density Upoproteins (LDL) and foreign antigens. Scavenger receptors types I and II are trimeric membrane proteins with each subunit containing a small N-terminal intracellular domain, a 5 transmembrane domain, a large extracellular domain, and a C-terminal cysteine-rich domain. The extracellular domain contains a short spacer domain, an α-heUcal coiled-coil domain, and a triple heUcal collagenous domain. These receptors have been shown to bind a spectrum of Ugands, including chemically modified Upoproteins and albumin, polyribonucleotides, polysaccharides, phosphoUpids, and asbestos (Matsumoto, A. et al. (1990) Proc. Natl. Acad. Sci. USA 87:9133-9137; Elomaa, O. et al. o (1995) Cell 80:603-609). The scavenger receptors are thought to play a key role in atherogenesis by mediating uptake of modified LDL in arterial walls, and in host defense by binding bacterial endotoxins, bacteria, and protozoa. T-Cell Receptors
T cells play a dual role in the immune system as effectors and regulators, coupUng antigen 5 recognition with the transmission of signals that induce cell death in infected cells and stimulate proUferation of other immune cells. Although a population of T cells can recognize a wide range of different antigens, an individual T cell can only recognize a single antigen and only when it is presented to the T cell receptor (TCR) as a peptide complexed with a major histocompatibiUty molecule (MHC) on the surface of an antigen presenting cell. The TCR on most T cells consists of immunoglobulin-like o integral membrane glycoproteins containing two polypeptide subunits, α and β, of similar molecular weight. Both TCR subunits have an extracellular domain containing both variable and constant regions, a transmembrane domain that traverses the membrane once, and a short intracellular domain (Saito, H. et al. (1984) Nature 309:757-762). The genes for the TCR subunits are constructed through somatic rearrangement of different gene segments. Interaction of antigen in the proper MHC context 5 with the TCR initiates signaling cascades that induce the proUferation, maturation, and function of cellular components of the immune system (Weiss, A. (1991) Annu. Rev. Genet. 25:487-510). Rearrangements in TCR genes and alterations in TCR expression have been noted in lymphomas, leukemias, autoimmune disorders, and immunodeficiency disorders (Asenberg, AC. et al. (1985) N. Engl. J. Med. 313:529-533; Weiss, supra). 0
Intracellular Signaling Molecules
SEQ ID NO:12, SEQ ID NO:13, SEQ ID NO:14, SEQ ID NO:15, SEQ ID NO:16, SEQ ID NO: 17, and SEQ ID NO: 18 encode, for example, intracellular signaling molecules.
Intracellular signaling is the general process by which cells respond to extracellular signals (hormones, neurotransmitters, growth and differentiation factors, etc.) through a cascade of biochemical reactions that begins with the binding of a signaling molecule to a cell membrane receptor and ends with the activation of an intracellular target molecule. Intermediate steps in the process involve the activation of various cytoplasmic proteins by phosphorylation via protein kinases, 5 and their deactivation by protein phosphatases, and the eventual translocation of some of these activated proteins to the cell nucleus where the transcription of specific genes is triggered. The intracellular signaling process regulates all types of cell functions including cell proliferation, cell differentiation, and gene transcription, and involves a diversity of molecules including protein kinases and phosphatases, and second messenger molecules, such as cyclic nucleotides, calcium-calmodulin, o inositol, and various mitogens, that regulate protein phosphorylation.
Protein Phosphorylation
Protein kinases and phosphatases play a key role in the intracellular signaling process by controlling the phosphorylation and activation of various signaling proteins. The high energy phosphate for this reaction is generally transferred from the adenosine triphosphate molecule (ATP) to 5 a particular protein by a protein kinase and removed from that protein by a protein phosphatase. Protein kinases are roughly divided into two groups: those that phosphorylate tyrosine residues (protein tyrosine kinases, PTK) and those that phosphorylate serine or threonine residues (serine/threonine kinases, STK). A few protein kinases have dual specificity for serine/threonine and tyrosine residues. Almost all kinases contain a conserved 250-300 amino acid catalytic domain o containing specific residues and sequence motifs characteristic of the kinase family (Hardie, G. and S.
Hanks (1995) The Protein Kinase Facts Books. Vol 1:7-20, Academic Press, San Diego CA).
STKs include the second messenger dependent protein kinases such as the cyclic- AMP dependent protein kinases (PKA), involved in mediating hormone-induced cellular responses; calcium-calmodulin (CaM) dependent protein kinases, involved in regulation of smooth muscle 5 contraction, glycogen breakdown, and neurofransmission; and the mitogen-activated protein kinases
(MAP) which mediate signal transduction from the cell surface to the nucleus via phosphorylation cascades. Altered PKA expression is implicated in a variety of disorders and diseases including cancer, thyroid disorders, diabetes, atherosclerosis, and cardiovascular disease (Isselbacher, KJ. et al. (1994) Harrison's Principles of Internal Medicine McGraw-Hill, New York NY, pp. 416-431, 1887). o PTKs are divided into transmembrane, receptor PTKs and nontransmembrane, non-receptor
PTKs. Transmembrane PTKs are receptors for most growth factors. Non-receptor PTKs lack transmembrane regions and, instead, form complexes with the intracellular regions of cell surface receptors. Receptors that function through non-receptor PTKs include those for cytokines and hormones (growth hormone and prolactin) and antigen-specific receptors on T and B lymphocytes. 5 Many of these PTKs were first identified as the products of mutant oncogenes in cancer cells in which their activation was no longer subject to normal cellular controls. In fact, about one third of the known oncogenes encode PTKs, and it is well known that cellular transformation (oncogenesis) is often accompanied by increased tyrosine phosphorylation activity (Charbonneau, H. and N.K. Tonks (1992) Annu. Rev. Cell Biol. 8:463-493). 5 An additional family of protein kinases previously thought to exist only in procaryotes is the histidine protein kinase family (HPK). HPKs bear little homology with mammalian STKs or PTKs but have distinctive sequence motifs of their own (Davie, J.R. et al. (1995) J. Biol. Chem. 270: 19861-19867). A histidine residue in the N-terminal half of the molecule (region I) is an autophosphorylation site. Three additional motifs located in the C-terminal half of the molecule 0 include an invariant asparagine residue in region II and two glycine-rich loops characteristic of nucleotide binding domains in regions III and IV. Recently a branched chain alpha-ketoacid dehydrogenase kinase has been found with characteristics of HPK in rat (Davie, supra).
Protein phosphatases regulate the effects of protein kinases by removing phosphate groups from molecules previously activated by kinases. The two principal categories of protein phosphatases 5 are the protein (serine/threonine) phosphatases (PPs) and the protein tyrosine phosphatases (PTPs). PPs dephosphorylate phosphoserine/threonine residues and are important regulators of many cAMP-mediated hormone responses (Cohen, P. (1989) Annu. Rev. Biochem. 58:453-508). PTPs reverse the effects of protein tyrosine kinases and play a significant role in cell cycle and cell signaling processes (Charbonneau, supra). As previously noted, many PTKs are encoded by o oncogenes, and oncogenesis is often accompanied by increased tyrosine phosphorylation activity. It is therefore possible that PTPs may prevent or reverse cell transformation and the growth of various cancers by controlling the levels of tyrosine phosphorylation in cells. This hypothesis is supported by studies showing that overexpression of PTPs can suppress transformation in cells, and that specific inhibition of PTPs can enhance cell transformation (Charbonneau, supra). 5 Phospholipid and Inositol-Phosphate Signaling
Inositol phospholipids (phosphoinositides) are involved in an intracellular signaling pathway that begins with binding of a signaling molecule to a G-protein linked receptor in the plasma membrane. This leads to the phosphorylation of phosphatidylinositol (PI) residues on the inner side of the plasma membrane to the biphosphate state (PIPj) by inositol kinases. Simultaneously, the G- o protein Unked receptor binding stimulates a trimeric G-protein which in turn activates a phosphoinositide-specific phospholipase C-β. PhosphoUpase C-β then cleaves PIP2 into two products, inositol triphosphate (IP3) and diacylglycerol. These two products act as mediators for separate signaling events. IP3 diffuses through the plasma membrane to induce calcium release from the endoplasmic reticulum (ER), while diacylglycerol remains in the membrane and helps activate 5 protein kinase C, an STK that phosphorylates selected proteins in the target cell. The calcium response initiated by IP3 is terminated by the dephosphorylation of IP3 by specific inositol phosphatases. Cellular responses that are mediated by this pathway are glycogen breakdown in the liver in response to vasopressin, smooth muscle contraction in response to acetylchoUne, and thrombin-induced platelet aggregation. 5 Cyclic Nucleotide SignaUng
Cyclic nucleotides (cAMP and cGMP) function as intracellular second messengers to transduce a variety of extracellular signals including hormones, light, and neurotransmitters. In particular, cyclic- AMP dependent protein kinases (PKA) are thought to account for all of the effects of cAMP in most mammaUan cells, including various hormone-induced cellular responses. Visual o excitation and the photottansmission of light signals in the eye is controlled by cyclic-GMP regulated,
Ca2+-specific channels. Because of the importance of cellular levels of cyclic nucleotides in mediating these various responses, regulating the synthesis and breakdown of cyclic nucleotides is an important matter. Thus adenylyl cyclase, which synthesizes cAMP from AMP, is activated to increase cAMP levels in muscle by binding of adrenaUne to β-andrenergic receptors, while activation 5 of guanylate cyclase and increased cGMP levels in photoreceptors leads to reopening of the
Ca2+-specific channels and recovery of the dark state in the eye. In contrast, hydrolysis of cyclic nucleotides by cAMP and cGMP-specific phosphodiesterases (PDEs) produces the opposite of these and other effects mediated by increased cycUc nucleotide levels. PDEs appear to be particularly important in the regulation of cyclic nucleotides, considering the diversity found in this family of o proteins. At least seven families of mammalian PDEs (PDE1 -7) have been identified based on substrate specificity and affinity, sensitivity to cofactors, and sensitivity to inhibitory drugs (Beavo, J.A. (1995) Physiological Reviews 75:725-48). PDE inhibitors have been found to be particularly useful in treating various clinical disorders. Rolipram, a specific inhibitor of PDE4, has been used in the treatment of depression, and similar inhibitors are undergoing evaluation as anti-inflammatory 5 agents. TheophyUine is a nonspecific PDE inhibitor used in the treatment of bronchial asthma and other respiratory diseases (Banner, K.H. and CP. Page (1995) Eur. Respir. J. 8:996-1000). G-Protein Signaling
Guanine nucleotide binding proteins (G-proteins) are critical mediators of signal transduction between a particular class of extracellular receptors, the G-protein coupled receptors (GPCR), and 0 intracellular second messengers such as cAMP and Ca2+. G-proteins are linked to the cytosoUc side of a GPCR such that activation of the GPCR by ligand binding stimulates binding of the G-protein to GTP, inducing an "active" state in the G-protein. In the active state, the G-protein acts as a signal to trigger other events in the cell such as the increase of cAMP levels or the release of Ca2+ into the cytosol from the ER, which, in turn, regulate phosphorylation and activation of other intracellular 5 proteins. Recycling of the G-protein to the inactive state involves hydrolysis of the bound GTP to GDP by a GTPase activity in the G-protein. (See Alberts, B. et al. (1994) Molecular Biology of the Cell, Garland Publishing, Inc., New York NY, pp.734-759.) Two structurally distinct classes of G- proteins are recognized: heterotrimeric G-proteins, consisting of three different subunits, and monomeric, low molecular weight (LMW), G-proteins consisting of a single polypeptide chain. 5 The three polypeptide subunits of heterotrimeric G-proteins are the cc, β, and γ subunits. The α subunit binds and hydrolyzes GTP. The β and γ subunits form a tight complex that anchors the protein to the inner side of the plasma membrane. The β subunits, also known as G-β proteins or β transducins, contain seven tandem repeats of the WD-repeat sequence motif, a motif found in many proteins with regulatory functions. Mutations and variant expression of β fransducin proteins are o linked with various disorders (Neer, E. J. et al. (1994) Nature 371 :297-300; Margottin, F. et al. (1998)
Mol. Cell 1:565-574).
LMW GTP-proteins are GTPases which regulate cell growth, cell cycle control, protein secretion, and intracellular vesicle interaction. They consist of single polypeptides which, Uke the α subunit of the heterotrimeric G-proteins, are able to bind and hydrolyze GTP, thus cycling between an 5 inactive and an active state. At least sixty members of the LMW G-protein superfamily have been identified and are currently grouped into the six subfamilies of ras, rho, arf, sari, ran, and rab. Activated ras genes were initially found in human cancers, and subsequent studies confirmed that ras function is critical in determining whether cells continue to grow or become differentiated. Other members of the LMW G-protein superfamily have roles in signal transduction that vary with the o function of the activated genes and the locations of the G-proteins.
Guanine nucleotide exchange factors regulate the activities of LMW G-proteins by determining whether GTP or GDP is bound. GTPase-activating protein (GAP) binds to GTP-ras and induces it to hydrolyze GTP to GDP. In contrast, guanine nucleotide releasing protein (GNRP) binds to GDP-ras and induces the release of GDP and the binding of GTP. 5 Other regulators of G-protein signaling (RGS) also exist that act primarily by negatively regulating the G-protein pathway by an unknown mechanism (Druey, KM. et al. (1996) Nature 379:742-746). Some 15 members of the RGS family have been identified. RGS family members are related structurally through similarities in an approximately 120 amino acid region termed the RGS domain and functionally by their abiUty to inhibit the interleukin (cytokine) induction of MAP kinase 0 in cultured mammalian 293T cells (Druey, supra). Calcium Signaling Molecules
Ca+2 is another second messenger molecule that is even more widely used as an intracellular mediator than cAMP. Two pathways exist by which Ca+2 can enter the cytosol in response to extracellular signals: One pathway acts primarily in nerve signal transduction where Ca+2 enters a nerve terminal through a voltage-gated Ca+2 channel. The second is a more ubiquitous pathway in which Ca+2 is released from the ER into the cytosol in response to binding of an extracellular signaling molecule to a receptor. Ca2+ directly activates regulatory enzymes, such as protein kinase C, which trigger signal transduction pathways. Ca2+ also binds to specific Ca2+-binding proteins (CBPs) 5 such as calmoduUn (CaM) which then activate multiple target proteins in the cell including enzymes, membrane transport pumps, and ion channels. CaM interactions are involved in a multitude of cellular processes including, but not limited to, gene regulation, DNA synthesis, cell cycle progression, mitosis, cytokinesis, cytoskeletal organization, muscle contraction, signal transduction, ion homeostasis, exocytosis, and metabolic regulation (Celio, M.R. et al. (1996) Guidebook to 0 Calcium-binding Proteins, Oxford University Press, Oxford, UK, pp. 15-20). Some CBPs can serve as a storage depot for Ca2+ in an inactive state. Calsequestrin is one such CBP that is expressed in isoforms specific to cardiac muscle and skeletal muscle. It is suggested that calsequestrin binds Ca2+ in a rapidly exchangeable state that is released during Ca2+ -signaling conditions (Celio, M.R. et al. (1996) Guidebook to Calcium-binding Proteins, Oxford University Press, New York NY, pp. 222- 5 224). Cvclins
Cell division is the fundamental process by which all Uving things grow and reproduce. In most organisms, the cell cycle consists of three principle steps; inteφhase, mitosis, and cytokinesis. Inteφhase, involves preparations for cell division, replication of the DNA and production of essential o proteins. In mitosis, the nuclear material is divided and separates to opposite sides of the cell.
Cytokinesis is the final division and fission of the cell cytoplasm to produce the daughter cells.
The entry and exit of a cell from mitosis is regulated by the synthesis and destruction of a family of activating proteins called cyclins. Cyclins act by binding to and activating a group of cycUn-dependent protein kinases (Cdks) which then phosphorylate and activate selected proteins 5 involved in the mitotic process. Several types of cyclins exist. (Ciechanover, A. (1994) Cell
79:13-21.) Two principle types are mitotic cycUn, or cyclin B, which controls entry of the cell into mitosis, and Gl cyclin, which controls events that drive the cell out of mitosis. Signal Complex Scaffolding Proteins
Ceretain proteins in intracellular signaling pathways serve to link or cluster other proteins o involved in the signaling cascade. A conserved protein domain called the PDZ domain has been identified in various membrane-associated signaling proteins. This domain has been implicated in receptor and ion channel clustering and in the targeting of multiprotein signaling complexes to specialized functional regions of the cytosoUc face of the plasma membrane. (For a review of PDZ domain-containing proteins, see Ponting, CP. et al. (1997) Bioessays 19:469-479.) A large proportion of PDZ domains are found in the eukaryotic MAGUK (membrane-associated guanylate kinase) protein family, members of which bind to the intracellular domains of receptors and channels. However, PDZ domains are also found in diverse membrane-localized proteins such as protein tyrosine phosphatases, serine/threonine kinases, G-protein cofactors, and synapse-associated proteins 5 such as syntrophins and neuronal nitric oxide synthase (nNOS). Generally, about one to three PDZ domains are found in a given protein, although up to nine PDZ domains have been identified in a single protein.
Membrane Transport Molecules o The plasma membrane acts as a barrier to most molecules. Transport between the cytoplasm and the extracellular environment, and between the cytoplasm and lumenal spaces of cellular organelles requires specific transport proteins. Each transport protein carries a particular class of molecule, such as ions, sugars, or amino acids, and often is specific to a certain molecular species of the class. A variety of human inherited diseases are caused by a mutation in a transport protein. For 5 example, cystinuria is an inherited disease that results from the inability to transport cystine, the disulfide-linked dimer of cysteine, from the urine into the blood. Accumulation of cystine in the urine leads to the formation of cystine stones in the kidneys.
Transport proteins are multi-pass transmembrane proteins, which either actively transport molecules across the membrane or passively allow them to cross. Active transport involves o directional pumping of a solute across the membrane, usually against an electrochemical gradient.
Active transport is tightly coupled to a source of metabolic energy, such as ATP hydrolysis or an elecfrochemically favorable ion gradient. Passive transport involves the movement of a solute down its electrochemical gradient. Transport proteins can be further classified as either carrier proteins or channel proteins. Carrier proteins, which can function in active or passive transport, bind to a specific 5 solute to be transported and undergo a conf ormational change which transfers the bound solute across the membrane. Channel proteins, which only function in passive transport, form hydrophilic pores across the membrane. When the pores open, specific solutes, such as inorganic ions, pass through the membrane and down the electrochemical gradient of the solute.
Carrier proteins which transport a single solute from one side of the membrane to the other 0 are called uniporters. In contrast, coupled transporters link the transfer of one solute with simultaneous or sequential transfer of a second solute, either in the same direction (symport) or in the opposite direction (antiport). For example, intestinal and kidney epithelium contains a variety of symporter systems driven by the sodium gradient that exists across the plasma membrane. Sodium moves into the cell down its electrochemical gradient and brings the solute into the cell with it. The 5 sodium gradient that provides the driving force for solute uptake is maintained by the ubiquitous Na7K+ ATPase. Sodium-coupled transporters include the mammaUan glucose transporter (SGLT1), iodide transporter (NIS), and multivitamin transporter (SMVT). Al three transporters have twelve putative transmembrane segments, extracellular glycosylation sites, and cytoplasmically-oriented N- and C-termini. NIS plays a crucial role in the evaluation, diagnosis, and treatment of various thyroid 5 pathologies because it is the molecular basis for radioiodide thyroid-imaging techniques and for specific targeting of radioisotopes to the thyroid gland (Levy, O. et al. (1997) Proc. Natl. Acad. Sci. USA 94:5568-5573). SMVT is expressed in the intestinal mucosa, kidney, and placenta, and is implicated in the transport of the water-soluble vitamins, e.g., biotin and pantothenate (Prasad, P.D. et al. (1998) J. Biol. Chem. 273:7501-7506). o Transporters play a major role in the regulation of pH, excretion of drugs, and the cellular
K7Na+ balance. Monocarboxylate anion transporters are proton-coupled symporters with a broad substrate specificity that includes L-lactate, pyruvate, and the ketone bodies acetate, acetoacetate, and beta-hydroxybutyrate. At least seven isoforms have been identified to date. The isoforms are predicted to have twelve transmembrane (TM) heUcal domains with a large intracellular loop between TM6 and 5 TM7, and play a critical role in maintaining intracellular pH by removing the protons that are produced stoichiometrically with lactate during glycolysis. The best characterized H(+)-monocarboxylate transporter is that of the erythrocyte membrane, which transports L-lactate and a wide range of other aUphatic monocarboxylates. Other cells possess H(+)-Unked monocarboxylate transporters with differing substrate and inhibitor selectivities. In particular, cardiac muscle and tumor cells have o transporters that differ in their Km values for certain substrates, including stereoselectivity for L- over
D-lactate, and in their sensitivity to inhibitors. There are Na(+)-monocarboxylate cotransporters on the luminal surface of intestinal and kidney epitheUa, which allow the uptake of lactate, pyruvate, and ketone bodies in these tissues. In addition, there are specific and selective transporters for organic cations and organic anions in organs including the kidney, intestine and liver. Organic anion 5 transporters are selective for hydrophobic, charged molecules with electron-attracting side groups.
Organic cation transporters, such as the ammonium transporter, mediate the secretion of a variety of drugs and endogenous metaboUtes, and contribute to the maintenance of intercellular pH. (Poole, R.C. and AP. Halestrap (1993) Am. J. Physiol. 264:C761-C782; Price, N.T. et al. (1998) Biochem. J. 329:321-328; and Martinelle, K. and I. Haggstrom (1993) J. Biotechnol. 30: 339-350.) o The largest and most diverse family of transport proteins known is the ATP-binding cassette
(ABC) transporters. As a family, ABC transporters can transport substances that differ markedly in chemical structure and size, ranging from small molecules such as ions, sugars, amino acids, peptides, and phosphoUpids, to Upopeptides, large proteins, and complex hydrophobic drugs. ABC proteins consist of four modules: two nucleotide-binding domains (NBD), which hydrolyze ATP to supply the energy required for transport, and two membrane-spanning domains (MSD), each containing six putative transmembrane segments. These four modules may be encoded by a single gene, as is the case for the cystic fibrosis transmembrane regulator (CFTR), or by separate genes. When encoded by separate genes, each gene product contains a single NBD and MSD. These "half-molecules" form 5 homo- and heterodimers, such as Tapl and Tap2, the endoplasmic reticulum-based major histocompatibiUty (MHC) peptide transport system. Several genetic diseases are attributed to defects in ABC transporters, such as the following diseases and their corresponding proteins: cystic fibrosis (CFTR, an ion channel), adrenoleukodystrophy (adrenoleukodysfrophy protein, ALDP), Zellweger syndrome (peroxisomal membrane protein-70, PMP70), and hyperinsuUnemic hypoglycemia 0 (sulfonylurea receptor, SUR). Overexpression of the multidrug resistance (MDR) protein, another
ABC transporter, in human cancer cells makes the cells resistant to a variety of cytotoxic drugs used in chemotherapy (TagUght, D. and S. MichaeUs (1998) Meth. Enzymol. 292:131-163).
Transport of fatty acids across the plasma membrane can occur by diffusion, a high capacity, low affinity process. However, under normal physiological conditions a significant fraction of fatty 5 acid transport appears to occur via a high affinity, low capacity protein-mediated transport process. Fatty acid transport protein (FATP), an integral membrane protein with four transmembrane segments, is expressed in tissues exhibiting high levels of plasma membrane fatty acid flux, such as muscle, heart, and adipose. Expression of FATP is upregulated in 3T3-L1 cells during adipose conversion, and expression in COS7 fibroblasts elevates uptake of long-chain fatty acids (Hui, T.Y. et al. (1998) J. o Biol. Chem. 273 :27420-27429). Ion Channels
The electrical potential of a cell is generated and maintained by controlUng the movement of ions across the plasma membrane. The movement of ions requires ion channels, which form an ion- selective pore within the membrane. There are two basic types of ion channels, ion transporters and 5 gated ion channels. Ion transporters utilize the energy obtained from ATP hydrolysis to actively transport an ion against the ion's concentration gradient. Gated ion channels allow passive flow of an ion down the ion's electrochemical gradient under restricted conditions. Together, these types of ion channels generate, maintain, and utitize an electrochemical gradient that is used in 1) electrical impulse conduction down the axon of a nerve cell, 2) transport of molecules into cells against concentration o gradients, 3) initiation of muscle contraction, and 4) endocrine cell secretion.
Ion transporters generate and maintain the resting electrical potential of a cell. UtiUzing the energy derived from ATP hydrolysis, they transport ions against the ion's concentration gradient. These transmembrane ATPases are divided into three famiUes. The phosphorylated (P) class ion transporters, including Na+-K+ ATPase, Ca2+-ATPase, and H+-ATPase, are activated by a phosphorylation event. P-class ion transporters are responsible for maintaining resting potential distributions such that cytosoUc concentrations of Na+ and Ca2+ are low and cytosoUc concentration of K+ is high. The vacuolar (V) class of ion transporters includes H+ pumps on intracellular organelles, such as lysosomes and Golgi. V-class ion transporters are responsible for generating the low pH within 5 the lumen of these organelles that is required for function. The coupUng factor (F) class consists of H+ pumps in the mitochondria. F-class ion transporters utiUze a proton gradient to generate ATP from ADP and inorganic phosphate (PJ.
The resting potential of the cell is utiUzed in many processes involving carrier proteins and gated ion channels. Carrier proteins utilize the resting potential to transport molecules into and out of 0 the cell. Amino acid and glucose transport into many cells is Unked to sodium ion co-transport
(symport) so that the movement of Na+ down an electrochemical gradient drives transport of the other molecule up a concentration gradient. Similarly, cardiac muscle links transfer of Ca2+ out of the cell with transport of Na+ into the cell (antiport).
Ion channels share common structural and mechanistic themes. The channel consists of four or 5 five subunits or protein monomers that are arranged like a barrel in the plasma membrane. Each subunit typically consists of six potential transmembrane segments (SI, S2, S3, S4, S5, and S6). The center of the barrel forms a pore lined by α-helices or β-strands. The side chains of the amino acid residues comprising the α-heUces or β-strands estabUsh the charge (cation or anion) selectivity of the channel. The degree of selectivity, or what specific ions are allowed to pass through the channel, o depends on the diameter of the narrowest part of the pore.
Gated ion channels control ion flow by regulating the opening and closing of pores. These channels are categorized according to the manner of regulating the gating function. Mechanically-gated channels open pores in response to mechanical stress, voltage-gated channels open pores in response to changes in membrane potential, and Ugand-gated channels open pores in the presence of a specific ion, 5 nucleotide, or neurotransmitter.
Voltage-gated Na+ and K+ channels are necessary for the function of electrically excitable cells, such as nerve and muscle cells. Action potentials, which lead to neurotransmitter release and muscle contraction, arise from large, transient changes in the permeabiUty of the membrane to Na+ and K+ ions. Depolarization of the membrane beyond the threshold level opens voltage-gated Na+ channels. Sodium o ions flow into the cell, further depolarizing the membrane and opening more voltage-gated Na+ channels, which propagates the depolarization down the length of the cell. Depolarization also opens voltage-gated potassium channels. Consequently, potassium ions flow outward, which leads to repolarization of the membrane. Voltage-gated channels utilize charged residues in the fourth transmembrane segment (S4) to sense voltage change. The open state lasts only about 1 millisecond, at which time the channel spontaneously converts into an inactive state that cannot be opened irrespective of the membrane potential. Inactivation is mediated by the channel's N-terminus, which acts as a plug that closes the pore. The transition from an inactive to a closed state requires a return to resting potential. Voltage-gated Na+ channels are heterotrimeric complexes composed of a 260 kDa pore forming α subunit that associates with two smaller auxiliary subunits, βl and β2. The β2 subunit is an integral membrane glycoprotein that contains an extracellular Ig domain, and its association with α and βl subunits correlates with increased functional expression of the channel, a change in its gating properties, and an increase in whole cell capacitance due to an increase in membrane surface area. (Isom, L.L. et al. (1995) Cell 83:433-442.)
Voltage-gated Ca2+ channels are involved in presynaptic neurotransmitter release, and heart and skeletal muscle contraction. The voltage-gated Ca2+ channels from skeletal muscle (L-type) and brain (N-type) have been purified, and though their functions differ dramatically, they have similar subunit compositions. The channels are composed of three subunits. The aλ subunit forms the membrane pore and voltage sensor, while the α2δ and β subunits modulate the voltage-dependence, gating properties, and the current amplitude of the channel. These subunits are encoded by at least six α1( one 028, and four β genes. A fourth subunit, γ, has been identified in skeletal muscle. (Walker, D. et al. (1998) J. Biol. Chem. 273:2361-2367; and Jay, S.D. et al. (1990) Science 248:490-492.) Chloride channels are necessary in endocrine secretion and in regulation of cytosoUc and organelle pH. In secretory epitheUal cells, Cl " enters the cell across a basolateral membrane through an Na+, K7C1" cotransporter, accumulating in the cell above its electrochemical equilibrium concentration. Secretion of Cl " from the apical surface, in response to hormonal stimulation, leads to flow of Na + and water into the secretory lumen. The cystic fibrosis transmembrane conductance regulator (CFTR) is a chloride channel encoded by the gene for cystic fibrosis, a common fatal genetic disorder in humans. Loss of CFTR function decreases transepitheUal water secretion and, as a result, the layers of mucus that coat the respiratory tree, pancreatic ducts, and intestine are dehydrated and difficult to clear. The resulting blockage of these sites leads to pancreatic insufficiency, "meconium ileus", and devastating "chronic obstructive pulmonary disease" (A-Awqati, Q. et al. (1992) J. Exp. Biol. 172:245-266).
Many intracellular organelles contain H+- ATPase pumps that generate transmembrane pH and electrochemical differences by moving protons from the cytosol to the organelle lumen. If the membrane of the organelle is permeable to other ions, then the electrochemical gradient can be abrogated without affecting the pH differential. In fact, removal of the electrochemical barrier allows more H + to be pumped across the membrane, increasing the pH differential. Cl " is the sole counterion of H + translocation in a number of organelles, including chromaffin granules, Golgi vesicles, lysosomes, and endosomes. Functions that require a low vacuolar pH include uptake of small molecules such as biogenic amines in chromaffin granules, processing of vacuolar constituents such as pro-hormones by proteolytic enzymes, and protein degradation in lysosomes (A-Awqati, supra).
Ligand-gated channels open their pores when an extracellular or intracellular mediator binds to 5 the channel. Neurotransmitter-gated channels are channels that open when a neurotransmitter binds to their extracellular domain. These channels exist in the postsynaptic membrane of nerve or muscle cells. There are two types of neurotransmitter-gated channels. Sodium channels open in response to excitatory neurotransmitters, such as acetylchoUne, glutamate, and serotonin. This opening causes an influx of Na+ and produces the initial locaUzed depolarization that activates the voltage-gated channels o and starts the action potential. Chloride channels open in response to inhibitory neurotransmitters, such as γ-aminobutyric acid (GABA) and glycine, leading to hypeφolarization of the membrane and the subsequent generation of an action potential.
Ligand-gated channels can be regulated by intracellular second messengers. Calcium-activated K+ channels are gated by internal calcium ions. In nerve cells, an influx of calcium during 5 depolarization opens K+ channels to modulate the magnitude of the action potential (Ishi, T.M. et al. (1997) Proc. Nail. Acad. Sci. USA 94:11651-11656). Cyclic micleotide-gated (CNG) channels are gated by cytosoUc cyclic nucleotides. The best examples of these are the cAMP-gated Na + channels involved in olfaction and the cGMP-gated cation channels involved in vision. Both systems involve ligand-mediated activation of a G-protein coupled receptor which then alters the level of cyclic o nucleotide within the cell.
Ion channels are expressed in a number of tissues where they are impUcated in a variety of processes. CNG channels, while abundantly expressed in photoreceptor and olfactory sensory cells, are also found in kidney, lung, pineal, retinal gangUon cells, testis, aorta, and brain. Calcium-activated K+ channels may be responsible for the vasodilatory effects of bradykinin in the kidney and for shunting 5 excess K+ from brain capillary endotheUal cells into the blood. They are also impUcated in repolarizing granulocytes after agonist-stimulated depolarization (Ishi, supra). Ion channels have been the target for many drug therapies. Neurotransmitter-gated channels have been targeted in therapies for treatment of insomnia, anxiety, depression, and schizophrenia. Voltage-gated channels have been targeted in therapies for arrhythmia, ischemic stroke, head trauma, and neurodegenerative disease (Taylor, CP. o and L.S. Narasimhan (1997) Adv. Pharmacol. 39:47-98). Disease Correlation
The etiology of numerous human diseases and disorders can be attributed to defects in the transport of molecules across membranes. Defects in the trafficking of membrane-bound transporters and ion channels are associated with several disorders, e.g. cystic fibrosis, glucose-galactose malabsoφtion syndrome, hypercholesterolemia, von Gierke disease, and certain forms of diabetes melUtus. Single-gene defect diseases resulting in an inabiUty to transport small molecules across membranes include, e.g., cystinuria, iminoglycinuria, Hartup disease, andFanconi disease (vant Hoff, W.G. (1996) Exp. Nephrol. 4:253-262; Talente, G.M. et al. (1994) Ann. Intern. Med. 120:218-226; 5 and Chillon, M. et al. (1995) New Engl. J. Med. 332:1475-1480).
Protein Modification and Maintenance Molecules
SEQ ID NO:34 encodes, for example, a protein modification and maintenance molecule. The cellular processes regulating modification and maintenance of protein molecules o coordinate their conformation, stabilization, and degradation. Each of these processes is mediated by key enzymes or proteins such as proteases, protease inhibitors, transferases, isomerases, and molecular chaperones.
Proteases
Proteases cleave proteins and peptides at the peptide bond that forms the backbone of the 5 peptide and protein chain. Proteolytic processing is essential to cell growth, differentiation, remodeling, and homeostasis as well as inflammation and immune response. Typical protein half- lives range from hours to a few days, so that within all living cells, precursor proteins are being cleaved to their active form, signal sequences proteolytically removed from targeted proteins, and aged or defective proteins degraded by proteolysis. Proteases function in bacterial, parasitic, and viral o invasion and replication within a host. Four principal categories of mammalian proteases have been identified based on active site structure, mechanism of action, and overall three-dimensional structure. (Beynon, R.J. and J.S. Bond (1994) Proteolytic Enzymes: A Practical Approach, Oxford University Press, New York NY, pp. 1-5).
The serine proteases (SPs) have a serine residue, usually within a conserved sequence, in an 5 active site composed of the serine, an aspartate, and a histidine residue. SPs include the digestive enzymes trypsin and chymotrypsin, components of the complement cascade and the blood-clotting cascade, and enzymes that control extracellular protein degradation. The main SP sub-families are trypases, which cleave after arginine or lysine; aspartases, which cleave after aspartate; chymases, which cleave after phenylalanine or leucine; metases, which cleavage after methionine; and serases o which cleave after serine. Enterokinase, the initiator of intestinal digestion, is a serine protease found in the intestinal brush border, where it cleaves the acidic propeptide from trypsinogen to yield active trypsin (Kitamoto, Y. et al. (1994) Proc. Natl. Acad. Sci. USA 91:7588-7592). Prolylcarboxypeptidase, a lysosomal serine peptidase that cleaves peptides such as angiotensin II and III and [des- Ag9] bradykinin, shares sequence homology with members of both the serine carboxypeptidase and prolylendopeptidase families (Tan, F. et al. (1993) J. Biol. Chem. 268:16631- 16638).
Cysteine proteases (CPs) have a cysteine as the major catalytic residue at an active site where catalysis proceeds via an intermediate thiol ester and is facilitated by adjacent histidine and aspartic 5 acid residues. CPs are involved in diverse cellular processes ranging from the processing of precursor proteins to intracellular degradation. Mammalian CPs include lysosomal cathepsins and cytosoUc calcium activated proteases, calpains. CPs are produced by monocytes, macrophages and other cells of the immune system which migrate to sites of inflammation and secrete molecules involved in tissue repair. Overabundance of these repair molecules plays a role in certain disorders. In o autoimmune diseases such as rheumatoid arthritis, secretion of the cysteine peptidase cathepsin C degrades collagen, laminin, elastin and other structural proteins found in the extracellular matrix of bones.
Aspartic proteases are members of the cathepsin family of lysosomal proteases and include pepsin A, gastricsin, chymosin, renin, and cathepsins D and E. Aspartic proteases have a pair of 5 aspartic acid residues in the active site, and are most active in the pH 2 - 3 range, in which one of the aspartate residues is ionized, the other un-ionized. Aspartic proteases include bacterial penicillopepsin, mammalian pepsin, renin, chymosin, and certain fungal proteases. Abnormal regulation and expression of cathepsins is evident in various inflammatory disease states. In cells isolated from inflamed synovia, the mRNA for stromelysin, cytokines, TIMP-1, cathepsin, gelatinase, o and other molecules is preferentially expressed. Expression of cathepsins L and D is elevated in synovial tissues from patients with rheumatoid arthritis and osteoarthritis. Cathepsin L expression may also contribute to the influx of mononuclear cells which exacerbates the destruction of the rheumatoid synovium. (Keyszer, G.M. (1995) Arthritis Rheum. 38:976-984.) The increased expression and differential regulation of the cathepsins are Unked to the metastatic potential of a variety of cancers and 5 as such are of therapeutic and prognostic interest (Chambers, AF. et al. (1993) Crit. Rev. Oncog. 4:95-114).
Metalloproteases have active sites that include two glutamic acid residues and one histidine residue that serve as binding sites for zinc. Carboxypeptidases A and B are the principal mammalian metalloproteases. Both are exoproteases of similar structure and active sites. Carboxypeptidase A, o like chymotrypsin, prefers C-terminal aromatic and aliphatic side chains of hydrophobic nature, whereas carboxypeptidase B is directed toward basic arginine and lysine residues. Glycoprotease (GCP), or O-sialoglycoprotein endopeptidase, is a metallopeptidase which specifically cleaves 0-sialoglycoproteins such as glycophorin A. Another metallopeptidase, placental leucine aminopeptidase (P-LAP) degrades several peptide hormones such as oxytocin and vasopressin, suggesting a role in maintaining homeostasis during pregnancy, and is expressed in several tissues (Rogi, T. et al. (1996) J. Biol. Chem. 271:56-61).
Ubiquitin proteases are associated with the ubiquitin conjugation system (UCS), a major pathway for the degradation of cellular proteins in eukaryotic cells and some bacteria. The UCS mediates the elimination of abnormal proteins and regulates the half-lives of important regulatory proteins that control cellular processes such as gene transcription and cell cycle progression. In the UCS pathway, proteins targeted for degradation are conjugated to a ubiquitin, a small heat stable protein. The ubiquitinated protein is then recognized and degraded by proteasome, a large, multisubunit proteolytic enzyme complex, and ubiquitin is released for reutilization by ubiquitin protease. The UCS is implicated in the degradation of mitotic cyclic kinases, oncoproteins, tumor suppressor genes such as p53, viral proteins, cell surface receptors associated with signal transduction, transcriptional regulators, and mutated or damaged proteins (Ciechanover, A. (1994) Cell 79:13-21). A murine proto-oncogene, Unp, encodes a nuclear ubiquitin protease whose overexpression leads to oncogenic transformation of NIH3T3 cells, and the human homolog of this gene is consistently elevated in small cell tumors and adenocarcinomas of the lung (Gray, D. A. (1995) Oncogene 10:2179-2183). Signal Peptidases
The mechanism for the translocation process into the endoplasmic reticulum (ER) involves the recognition of an N-terminal signal peptide on the elongating protein. The signal peptide directs the protein and attached ribosome to a receptor on the ER membrane. The polypeptide chain passes through a pore in the ER membrane into the lumen while the N-terminal signal peptide remains attached at the membrane surface. The process is completed when signal peptidase located inside the ER cleaves the signal peptide from the protein and releases the protein into the lumen. Protease Inhibitors Protease inhibitors and other regulators of protease activity control the activity and effects of proteases. Protease inhibitors have been shown to control pathogenesis in animal models of proteolytic disorders (Muφhy, G. (1991) Agents Actions Suppl. 35:69-76). Low levels of the cystatins, low molecular weight inhibitors of the cysteine proteases, correlate with malignant progression of tumors. (Calkins, C. et al (1995) Biol. Biochem. Hoppe Seyler 376:71-80). Seφins are inhibitors of mammalian plasma serine proteases. Many seφins serve to regulate the blood clotting cascade and/or the complement cascade in mammals. Sp32 is a positive regulator of the mammalian acrosomal protease, acrosin, that binds the proenzyme, proacrosin, and thereby aides in packaging the enzyme into the acrosomal matrix (Baba, T. et al. (1994) J. Biol. Chem. 269:10133- 10140). The Kunitz family of serine protease inhibitors are characterized by one or more "Kunitz domains" containing a series of cysteine residues that are regularly spaced over approximately 50 amino acid residues and form three intrachain disulfide bonds. Members of this family include aprotinin, tissue factor pathway inhibitor (TFPI-1 and TFPI-2), inter-α-trypsin inhibitor, and bikunin. (Marlor, C.W. et al. (1997) J. Biol. Chem. 272:12202-12208.) Members of this family are potent inhibitors (in the nanomolar range) against serine proteases such as kalUkrein and plasmin. Aprotinin 5 has cUnical utiUty in reduction of perioperative blood loss.
A major portion of all proteins synthesized in eukaryotic cells are synthesized on the cytosolic surface of the endoplasmic reticulum (ER). Before these immature proteins are distributed to other organelles in the cell or are secreted, they must be transported into the interior lumen of the ER where post-translational modifications are performed. These modifications include protein folding o and the formation of disulfide bonds, and N-linked glycosylations.
Protein Isomerases
Protein folding in the ER is aided by two principal types of protein isomerases, protein disulfide isomerase (PDI), and peptidyl-prolyl isomerase (PPI)- PDI catalyzes the oxidation of free sulfhydryl groups in cysteine residues to form intramolecular disulfide bonds in proteins. PPI, an 5 enzyme that catalyzes the isomerization of certain proline imidic bonds in oligopeptides and proteins, is considered to govern one of the rate limiting steps in the folding of many proteins to their final functional conformation. The cyclophilins represent a major class of PPI that was originally identified as the major receptor for the immunosuppressive drug cyclosporin A (Handschumacher, RE. et al. (1984) Science 226: 544-547). 0 Protein Glvcosylation
The glycosylation of most soluble secreted and membrane-bound proteins by oligosaccharides linked to asparagine residues in proteins is also performed in the ER. This reaction is catalyzed by a membrane-bound enzyme, oUgosaccharyl transferase. Although the exact puφose of this "N-linked" glycosylation is unknown, the presence of oligosaccharides tends to make a 5 glycoprotein resistant to protease digestion. In addition, oligosaccharides attached to cell-surface proteins called selectins are known to function in cell-cell adhesion processes (Alberts, B. et al. (1994) Molecular Biology of the Cell, Garland Publishing Co., New York NY, p.608). "O-linked" glycosylation of proteins also occurs in the ER by the addition of N-acetylgalactosamine to the hydroxyl group of a serine or threonine residue followed by the sequential addition of other sugar o residues to the first. This process is catalysed by a series of glycosyltransferases each specific for a particular donor sugar nucleotide and acceptor molecule (Lodish, H. et al. (1995) Molecular Cell Biology, W.H. Freeman and Co., New York NY, pp.700-708). In many cases, both N- and O-linked oligosaccharides appear to be required for the secretion of proteins or the movement of plasma membrane glycoproteins to the cell surface. An additional glycosylation mechanism operates in the ER specifically to target lysosomal enzymes to lysosomes and prevent their secretion. Lysosomal enzymes in the ER receive an N-Unked oligosaccharide, like plasma membrane and secreted proteins, but are then phosphorylated on one or two mannose residues. The phosphorylation of mannose residues occurs in two steps, the first step 5 being the addition of an N-acetylglucosamine phosphate residue by N-acetylglucosamine phosphotransferase, and the second the removal of the N-acetylglucosamine group by phosphodiesterase. The phosphorylated mannose residue then targets the lysosomal enzyme to a mannose 6-phosphate receptor which transports it to a lysosome vesicle (Lodish, supra, pp. 708-711). Chaperones o Molecular chaperones are proteins that aid in the proper folding of immature proteins and refolding of improperly folded ones, the assembly of protein subunits, and in the transport of unfolded proteins across membranes. Chaperones are also called heat-shock proteins (hsp) because of their tendency to be expressed in dramatically increased amounts following brief exposure of cells to elevated temperatures. This latter property most likely reflects their need in the refolding of proteins 5 that have become denatured by the high temperatures. Chaperones may be divided into several classes according to their location, function, and molecular weight, and include hsp60, TCP1, hsp70, hsp40 (also called DnaJ), and hsp90. For example, hsp90 binds to steroid hormone receptors, represses transcription in the absence of the Ugand, and provides proper folding of the ligand-binding domain of the receptor in the presence of the hormone (Burston, S.G. and A.R. Clarke (1995) Essays 0 Biochem. 29:125-136). Hsp60 and hsp70 chaperones aid in the transport and folding of newly synthesized proteins. Hsp70 acts early in protein folding, binding a newly synthesized protein before it leaves the ribosome and transporting the protein to the mitochondria or ER before releasing the folded protein. Hsp60, along with hsp 10, binds misfolded proteins and gives them the opportunity to refold correctly. Al chaperones share an affinity for hydrophobic patches on incompletely folded 5 proteins and the ability to hydrolyze ATP. The energy of ATP hydrolysis is used to release the hsp- bound protein in its properly folded state (Aberts, supra, pp 214, 571-572).
Nucleic Acid Synthesis and Modification Molecules
SEQ ID NO:35 and SEQ ID NO:36 encode, for example, nucleic acid synthesis and o modification molecules .
Polymerases
DNA and RNA replication are critical processes for cell replication and function. DNA and RNA replication are mediated by the enzymes DNA and RNA polymerase, respectively, by a "templating" process in which the nucleotide sequence of a DNA or RNA strand is copied by 5 complementary base-pairing into a complementary nucleic acid sequence of either DNA or RNA. However, there are fundamental differences between the two processes.
DNA polymerase catalyzes the stepwise addition of a deoxyribonucleotide to the 3' -OH end of a polynucleotide strand (the primer strand) that is paired to a second (template) strand. The new DNA strand therefore grows in the 5' to 3' direction (Alberts, B. et al. (1994)The Molecular Biology 5 of the Cell, Garland Publishing Inc., New York NY, pp. 251-254). The substrates for the polymerization reaction are the corresponding deoxynucleotide triphosphates which must base-pair with the correct nucleotide on the template strand in order to be recognized by the polymerase. Because DNA exists as a double-stranded helix, each of the two strands may serve as a template for the formation of a new complementary strand. Each of the two daughter cells of the dividing cell 0 therefore inherits a new DNA double helix containing one old and one new strand. Thus, DNA is said to be replicated "semiconservatively" by DNA polymerase. In addition to the synthesis of new DNA, DNA polymerase is also involved in the repair of damaged DNA as discussed below under "Ligases."
In contrast to DNA polymerase, RNA polymerase uses a DNA template strand to "transcribe" 5 DNA into RNA using ribonucleotide triphosphates as substrates. Like DNA polymerization, RNA polymerization proceeds in a 5' to 3' direction by addition of a ribonucleoside monophosphate to the 3' -OH end of a growing RNA chain. DNA transcription generates messenger RNAs (mRNA) that carry information for protein synthesis, as well as the transfer, ribosomal, and other RNAs that have structural or catalytic functions. In eukaryotes, three discrete RNA polymerases synthesize the three o different types of RNA (Aberts, supra, pp. 367-368). RNA polymerase I makes the large ribosomal RNAs, RNA polymerase II makes the mRNAs that will be translated into proteins, and RNA polymerase III makes a variety of small, stable RNAs, including 5S ribosomal RNA and the transfer RNAs (tRNA). In all cases, RNA synthesis is initiated by binding of the RNA polymerase to a promoter region on the DNA and synthesis begins at a start site within the promoter. Synthesis is 5 completed at a broad, general stop or termination region in the DNA where both the polymerase and the completed RNA chain are released. Ligases
DNA repair is the process by which accidental base changes, such as those produced by oxidative damage, hydrolytic attack, or uncontrolled methylation of DNA are corrected before o replication or transcription of the DNA can occur. Because of the efficiency of the DNA repair process, fewer than one in one thousand accidental base changes causes a mutation (Alberts, supra, pp. 245-249). The three steps common to most types of DNA repair are (1) excision of the damaged or altered base or nucleotide by DNA nucleases, leaving a gap; (2) insertion of the correct nucleotide in this gap by DNA polymerase using the complementary strand as the template; and (3) sealing the 5 break left between the inserted nucleotide(s) and the existing DNA strand by DNA ligase. In the last reaction, DNA ligase uses the energy from ATP hydrolysis to activate the 5 ' end of the broken phosphodiester bond before forming the new bond with the 3'-OH of the DNA strand. In Bloom's syndrome, an inherited human disease, individuals are partially deficient in DNA ligation and consequently have an increased incidence of cancer (Alberts, supra, p. 247). 5 Nucleases
Nucleases comprise both enzymes that hydrolyze DNA (DNase) and RNA (RNase). They serve different puφoses in nucleic acid metaboUsm. Nucleases hydrolyze the phosphodiester bonds between adjacent nucleotides either at internal positions (endonucleases) or at the terminal 3' or 5' nucleotide positions (exonucleases). A DNA exonuclease activity in DNA polymerase, for example, o serves to remove improperly paired nucleotides attached to the 3'-OH end of the growing DNA strand by the polymerase and thereby serves a "proofreading" function. As mentioned above, DNA endonuclease activity is involved in the excision step of the DNA repair process.
RNases also serve a variety of functions. For example, RNase P is a ribonucleoprotein enzyme which cleaves the 5' end of pre-tRNAs as part of their maturation process. RNase H digests 5 the RNA strand of an RN A/DNA hybrid. Such hybrids occur in cells invaded by retroviruses, and RNase H is an important enzyme in the retroviral replication cycle. Pancreatic RNase secreted by the pancreas into the intestine hydrolyzes RNA present in ingested foods. RNase activity in serum and cell extracts is elevated in a variety of cancers and infectious diseases (Schein, CH. (1997) Nat. Biotechnol. 15:529-536). Regulation of RNase activity is being investigated as a means to control o tumor angiogenesis, allergic reactions, viral infection and replication, and fungal infections.
Methyl ases
Methylation of specific nucleotides occurs in both DNA and RNA, and serves different functions in the two macromolecules. Methylation of cytosine residues to form 5-methyl cytosine in DNA occurs specifically at CG sequences which are base-paired with one another in the DNA double- 5 helix. This pattern of methylation is passed from generation to generation during DNA replication by an enzyme called "maintenance methylase" that acts preferentially on those CG sequences that are base-paired with a CG sequence that is already methylated. Such methylation appears to distinguish active from inactive genes by preventing the binding of regulatory proteins that "turn on" the gene, but permit the binding of proteins that inactivate the gene (Alberts, supra, pp. 448-451). In RNA o metabolism, "tRNA methylase" produces one of several nucleotide modifications in tRNA that affect the conformation and base-pairing of the molecule and facilitate the recognition of the appropriate mRNA codons by specific tRNAs. The primary methylation pattern is the dimethylation of guanine residues to form N,N-dimethyl guanine. HeUcases and Single-Stranded Binding Proteins 5 HeUcases are enzymes that destabiUze and unwind double helix structures in both DNA and RNA. Since DNA replication occurs more or less simultaneously on both strands, the two strands must first separate to generate a replication "fork" for DNA polymerase to act on. Two types of replication proteins contribute to this process, DNA heUcases and single-stranded binding proteins. DNA heUcases hydrolyze ATP and use the energy of hydrolysis to separate the DNA strands. Single- 5 stranded binding proteins (SSBs) then bind to the exposed DNA strands without covering the bases, thereby temporarily stabilizing them for templating by the DNA polymerase (Alberts, supra, pp. 255- 256).
RNA helicases also alter and regulate RNA conformation and secondary structure. Like the DNA helicases, RNA helicases utilize energy derived from ATP hydrolysis to destabilize and unwind l o RNA duplexes. The most well-characterized and ubiquitous family of RNA helicases is the DEAD- box family, so named for the conserved B-type ATP-binding motif which is diagnostic of proteins in this family. Over 40 DEAD-box helicases have been identified in organisms as diverse as bacteria, insects, yeast, amphibians, mammals, and plants. DEAD-box helicases function in diverse processes such as translation initiation, splicing, ribosome assembly, and RNA editing, transport, and stability.
15 Some DEAD-box helicases play tissue- and stage-specific roles in spermatogenesis and embryogenesis. Overexpression of the DEAD-box 1 protein (DDX1) may play a role in the progression of neuroblastoma (Nb) and retinoblastoma (Rb) tumors (Godbout, R. et al. (1998) J. Biol. Chem. 273:21161-21168). These observations suggest that DDX1 may promote or enhance tumor progression by altering the normal secondary structure and expression levels of RNA in cancer cells.
2 o Other DEAD-box helicases have been implicated either directly or indirectly in tumorigenesis
(Discussed in Godbout, supra). For example, murine p68 is mutated in ultraviolet light-induced tumors, and human DDX6 is located at a chromosomal breakpoint associated with B-cell lymphoma. Similarly, a chimeric protein comprised of DDX10 and NUP98, a nucleoporin protein, may be involved in the pathogenesis of certain myeloid malignancies. 25 Topoisomerases
Besides the need to separate DNA strands prior to replication, the two strands must be "unwound" from one another prior to their separation by DNA helicases. This function is performed by proteins known as DNA topoisomerases. DNA topoisomerase effectively acts as a reversible nuclease that hydrolyzes a phosphodiesterase bond in a DNA strand, permitting the two strands to
3 o rotate freely about one another to remove the strain of the helix, and then rejoins the original phosphodiester bond between the two strands. Two types of DNA topoisomerase exist, types I and II. DNA Topoisomerase I causes a single-strand break in a DNA helix to allow the rotation of the two strands of the helix about the remaining phosphodiester bond in the opposite strand. DNA topoisomerase II causes a transient break in both strands of a DNA helix where two double heUces 35 cross over one another. This type of topoisomerase can efficiently separate two interlocked DNA circles (Aberts, supra, pp.260-262). Type II topoisomerases are largely confined to proliferating cells in eukaryotes, such as cancer cells. For this reason they are targets for anticancer drugs. Topoisomerase II has been implicated in multi-drug resistance (MDR) as it appears to aid in the repair of DNA damage inflicted by DNA binding agents such as doxorubicin and vincristine. Recombinases
Genetic recombination is the process of rearranging DNA sequences within an organism's genome to provide genetic variation for the organism in response to changes in the environment. DNA recombination allows variation in the particular combination of genes present in an individual's genome, as well as the timing and level of expression of these genes (see Aberts, supra, pp. 263- 273). Two broad classes of genetic recombination are commonly recognized, general recombination and site-specific recombination. General recombination involves genetic exchange between any homologous pair of DNA sequences usually located on two copies of the same chromosome. The process is aided by enzymes called recombinases that "nick" one strand of a DNA duplex more or less randomly and permit exchange with the complementary strand of another duplex. The process does not normally change the arrangement of genes on a chromosome. In site-specific recombination, the recombinase recognizes specific nucleotide sequences present in one or both of the recombining molecules. Base-pairing is not involved in this form of recombination and therefore does not require DNA homology between the recombining molecules. Unlike general recombination, this form of recombination can alter the relative positions of nucleotide sequences in chromosomes. Splicing Factors
Various proteins are necessary for processing of transcribed RNAs in the nucleus. Pre- mRNA processing steps include capping at the 5' end with methylguanosine, polyadenylating the 3' end, and splicing to remove introns. The primary RNA transcript from DNA is a faithful copy of the gene containing both exon and intron sequences, and the latter sequences must be cut out of the RNA transcript to produce an mRNA that codes for a protein. This "splicing" of the mRNA sequence takes place in the nucleus with the aid of a large, multicomponent ribonucleoprotein complex known as a sphceosome. The spUceosomal complex is composed of five small nuclear ribonucleoprotein particles (snRNPs) designated Ul, U2, U4, U5, and U6, and a number of additional proteins. Each snRNP contains a single species of snRNA and about ten proteins. The RNA components of some snRNPs recognize and base pair with intron consensus sequences. The protein components mediate sphceosome assembly and the splicing reaction. Autoantibodies to snRNP proteins are found in the blood of patients with systemic lupus erythematosus (Stryer, L. (1995) Biochemistry, W.H. Freeman and Company, New York NY, p. 863).
Adhesion Molecules The surface of a cell is rich in transmembrane proteoglycans, glycoproteins, glycoUpids, and receptors. These macromolecules mediate adhesion with other cells and with components of the extracellular matrix (ECM). The interaction of the cell with its surroundings profoundly influences cell shape, strength, flexibility, motiUty, and adhesion. These dynamic properties are intimately 5 associated with signal transduction pathways controlUng cell proliferation and differentiation, tissue construction, and embryonic development. Cadherins
Cadherins comprise a family of calcium-dependent glycoproteins that function in mediating cell-cell adhesion in virtually all soUd tissues of multicellular organisms. These proteins share o multiple repeats of a cadherin-specific motif, and the repeats form the folding units of the cadherin extracellular domain. Cadherin molecules cooperate to form focal contacts, or adhesion plaques, between adjacent epithelial cells. The cadherin family includes the classical cadherins and protocadherins. Classical cadherins include the E-cadherin, N-cadherin, and P-cadherin subfamilies. E-cadherin is present on many types of epithelial cells and is especially important for embryonic 5 development. N-cadherin is present on nerve, muscle, and lens cells and is also critical for embryonic development. P-cadherin is present on cells of the placenta and epidermis. Recent studies report that protocadherins are involved in a variety of cell-cell interactions (Suzuki, S.T. (1996) J. Cell Sci. 109:2609-2611). The intracellular anchorage of cadherins is regulated by their dynamic association with catenins, a family of cytoplasmic signal transduction proteins associated with the actin o cytoskeleton. The anchorage of cadherins to the actin cytoskeleton appears to be regulated by protein tyrosine phosphorylation, and the cadherins are the target of phosphorylation-induced junctional disassembly (Aberle, H. et al. (1996) J. Cell. Biochem. 61:514-523).
Integrins
Integrins are ubiquitous transmembrane adhesion molecules that link the ECM to the internal 5 cytoskeleton. Integrins are composed of two noncovalently associated transmembrane glycoprotein subunits called and β. Integrins function as receptors that play a role in signal transduction. For example, binding of integrin to its extracellular ligand may stimulate changes in intracellular calcium levels or protein kinase activity (Sjaastad, M.D. and W.J. Nelson (1997) BioEssays 19:47-55). At least ten cell surface receptors of the integrin family recognize the ECM component fibronectin, o which is involved in many different biological processes including cell migration and embryogenesis
(Johansson, S. et al. (1997) Front. Biosci. 2:D126-D146). Lectins
Lectins comprise a ubiquitous family of extracellular glycoproteins which bind cell surface carbohydrates specifically and reversibly, resulting in the agglutination of cells (reviewed in 5 Drickamer, K. and M.E. Taylor (1993) Annu. Rev. Cell Biol. 9:237-264). This function is particularly important for activation of the immune response. Lectins mediate the agglutination and mitogenic stimulation of lymphocytes at sites of inflammation (Lasky, L.A. (1991) J. Cell. Biochem. 45:139-146; Paietta, E. et al. (1989) J. Immunol. 143:2850-2857).
Lectins are further classified into subfamiUes based on carbohydrate-binding specificity and 5 other criteria. The galectin subfamily, in particular, includes lectins that bind β-galactoside carbohydrate moieties in a thiol-dependent manner (reviewed in Hadari, Y.R. et al. (1998) J. Biol. Chem. 270:3447-3453). Galectins are widely expressed and developmentally regulated. Because all galectins lack an N-terminal signal peptide, it is suggested that galectins are externalized through an atypical secretory mechanism. Two classes of galectins have been defined based on molecular weight 0 and oUgomerization properties. Small galectins form homodimers and are about 14 to 16 kilodaltons in mass, while large galectins are monomeric and about 29-37 kilodaltons.
Galectins contain a characteristic carbohydrate recognition domain (CRD). The CRD is about 140 amino acids and contains several stretches of about 1 - 10 amino acids which are highly conserved among all galectins. A particular 6-amino acid motif within the CRD contains conserved 5 tryptophan and arginine residues which are critical for carbohydrate binding. The CRD of some galectins also contains cysteine residues which may be important for disulfide bond formation. Secondary structure predictions indicate that the CRD forms several β-sheets.
Galectins play a number of roles in diseases and conditions associated with cell-cell and cell- matrix interactions. For example, certain galectins associate with sites of inflammation and bind to o cell surface immunoglobulin E molecules. In addition, galectins may play an important role in cancer metastasis. Galectin overexpression is correlated with the metastatic potential of cancers in humans and mice. Moreover, anti-galectin antibodies inhibit processes associated with cell transformation, such as cell aggregation and anchorage-independent growth (See, for example, Su, Z.-Z. et al. (1996) Proc. Natl. Acad. Sci. USA 93:7252-7257). 5 Selectins
Selectins, or LEC-CAMs, comprise a specialized lectin subfamily involved primarily in inflammation and leukocyte adhesion (Reviewed in Lasky, supra). Selectins mediate the recruitment of leukocytes from the circulation to sites of acute inflammation and are expressed on the surface of vascular endothelial cells in response to cytokine signaling. Selectins bind to specific ligands on the o leukocyte cell membrane and enable the leukocyte to adhere to and migrate along the endothelial surface. Binding of selectin to its ligand leads to polarized rearrangement of the actin cytoskeleton and stimulates signal transduction within the leukocyte (Brenner, B. et al. (1997) Biochem. Biophys. Res. Commun. 231:802-807; Hidari, KI. et al. (1997) J. Biol. Chem. 272:28750-28756). Members of the selectin family possess three characteristic motifs: a lectin or carbohydrate recognition domain; 5 an epidermal growth factor-like domain; and a variable number of short consensus repeats (scr or "sushi" repeats) which are also present in complement regulatory proteins. The selectins include lymphocyte adhesion molecule- 1 (Lam-1 or L-selectin), endothelial leukocyte adhesion molecule-1 (ELAM-1 or E-selectin), and granule membrane protein- 140 (GMP-140 or P-selectin) (Johnston, G.I. et al. (1989) Cell 56:1033-1044).
Antigen Recognition Molecules
SEQ ID NO:37 encodes, for example, an antigen recognition molecule. Al vertebrates have developed sophisticated and complex immune systems that provide protection from viral, bacterial, fungal, and parasitic infections. A key feature of the immune system is its ability to distinguish foreign molecules, or antigens, from "self molecules. This ability is mediated primarily by secreted and transmembrane proteins expressed by leukocytes (white blood cells) such as lymphocytes, granulocytes, and monocytes. Most of these proteins belong to the immunoglobulin (Ig) superfamily, members of which contain one or more repeats of a conserved structural domain. This Ig domain is comprised of antiparallel β sheets joined by a disulfide bond in an arrangement called the Ig fold. Members of the Ig superfamily include T-cell receptors, major histocompatibiUty (MHC) proteins, antibodies, and immune cell-specific surface markers such as CD4, CD8, and CD28.
MHC proteins are cell surface markers that bind to and present foreign antigens to T cells. MHC molecules are classified as either class I or class II. Class I MHC molecules (MHC I) are expressed on the surface of almost all cells and are involved in the presentation of antigen to cytotoxic T cells. For example, a cell infected with virus will degrade intracellular viral proteins and express the protein fragments bound to MHC I molecules on the cell surface. The MHC I/antigen complex is recognized by cytotoxic T-cells which destroy the infected cell and the virus within. Class II MHC molecules are expressed primarily on specialized antigen-presenting cells of the immune system, such as B-cells and macrophages. These cells ingest foreign proteins from the extracellular fluid and express MHC Il/antigen complex on the cell surface. This complex activates helper T-cells, which then secrete cytokines and other factors that stimulate the immune response. MHC molecules also play an important role in organ rejection following transplantation. Rejection occurs when the recipient's T-cells respond to foreign MHC molecules on the transplanted organ in the same way as to self MHC molecules bound to foreign antigen. (Reviewed in Aberts, B. et al. (1994) Molecular Biology of the Cell. Garland PubUshing, New York NY, pp. 1229-1246.)
Aitibodies, or immunoglobulins, are either expressed on the surface of B-cells or secreted by B-cells into the circulation. Aitibodies bind and neutralize foreign antigens in the blood and other extracellular fluids. The prototypical antibody is a tetramer consisting of two identical heavy polypeptide chains (H-chains) and two identical light polypeptide chains (L-chains) interlinked by disulfide bonds. This arrangement confers the characteristic Y-shape to antibody molecules. Antibodies are classified based on their H-chain composition. The five antibody classes, IgA, IgD, IgE, IgG and IgM, are defined by the α, δ, e, γ, and μ H-chain types. There are two types of L- chains, K and λ, either of which may associate as a pair with any H-chain pair. IgG, the most 5 common class of antibody found in the circulation, is tetrameric, while the other classes of antibodies are generally variants or multimers of this basic structure.
H-chains and L-chains each contain an N-terminal variable region and a C-terminal constant region. The constant region consists of about 110 amino acids in L-chains and about 330 or 440 amino acids in H-chains. The amino acid sequence of the constant region is nearly identical among 0 H- or L-chains of a particular class. The variable region consists of about 110 amino acids in both H- and L-chains. However, the amino acid sequence of the variable region differs among H- or L-chains of a particular class. Within each H- or L-chain variable region are three hypervariable regions of extensive sequence diversity, each consisting of about 5 to 10 amino acids. In the antibody molecule, the H- and L-chain hypervariable regions come together to form the antigen recognition site. 5 (Reviewed in Aberts, supra, pp. 1206-1213 and 1216-1217.)
Both H-chains and L-chains contain repeated Ig domains. For example, a typical H-chain contains four Ig domains, three of which occur within the constant region and one of which occurs within the variable region and contributes to the formation of the antigen recognition site. Likewise, a typical L-chain contains two Ig domains, one of which occurs within the constant region and one of o which occurs within the variable region.
The immune system is capable of recognizing and responding to any foreign molecule that enters the body. Therefore, the immune system must be armed with a full repertoire of antibodies against all potential antigens. Such antibody diversity is generated by somatic rearrangement of gene segments encoding variable and constant regions. These gene segments are joined together by site- 5 specific recombination which occurs between highly conserved DNA sequences that flank each gene segment. Because there are hundreds of different gene segments, millions of unique genes can be generated combinatorially. In addition, imprecise joining of these segments and an unusually high rate of somatic mutation within these segments further contribute to the generation of a diverse antibody population. o T-cell receptors are both structurally and functionally related to antibodies. (Reviewed in
Aberts, supra, pp. 1228-1229.) T-cell receptors are cell surface proteins that bind foreign antigens and mediate diverse aspects of the immune response. A typical T-cell receptor is a heterodimer comprised of two disulfide-Unked polypeptide chains called α and β. Each chain is about 280 amino acids in length and contains one variable region and one constant region. Each variable or constant region folds 5 into an Ig domain. The variable regions from the α and β chains come together in the heterodimer to form the antigen recognition site. T-cell receptor diversity is generated by somatic rearrangement of gene segments encoding the α and β chains. T-cell receptors recognize small peptide antigens that are expressed on the surface of antigen-presenting cells and pathogen-infected cells. These peptide antigens are presented on the cell surface in association with major histocompatibiUty proteins which provide the 5 proper context for antigen recognition.
Secreted and Extracellular Matrix Molecules
SEQ ID NO:38 and SEQ ID NO:39 encode, for example, secreted/extracellular matrix molecules. 0 Protein secretion is essential for cellular function. Protein secretion is mediated by a signal peptide located at the amino terminus of the protein to be secreted. The signal peptide is comprised of about ten to twenty hydrophobic amino acids which target the nascent protein from the ribosome to the endoplasmic reticulum (ER). Proteins targeted to the ER may either proceed through the secretory pathway or remain in any of the secretory organelles such as theER, Golgi apparatus, or lysosomes. 5 Proteins that transit through the secretory pathway are either secreted into the extracellular space or retained in the plasma membrane. Secreted proteins are often synthesized as inactive precursors that are activated by post-translational processing events during transit through the secretory pathway. Such events include glycosylation, proteolysis, and removal of the signal peptide by a signal peptidase. Other events that may occur during protein transport include chaperone-dependent unfolding and o folding of the nascent protein and interaction of the protein with a receptor or pore complex. Examples of secreted proteins with amino terminal signal peptides include receptors, extracellular matrix molecules, cytokines, hormones, growth and differentiation factors, neuropeptides, vasomediators, ion channels, transporters/pumps, and proteases. (Reviewed in Aberts, B. et al. (1994) Molecular Biology of The Cell, Garland PubUshing, New York NY, pp. 557-560, 582-592.) 5 The extracellular matrix (ECM) is a complex network of glycoproteins, polysaccharides, proteoglycans, and other macromolecules that are secreted from the cell into the extracellular space. The ECM remains in close association with the cell surface and provides a supportive meshwork that profoundly influences cell shape, motility, strength, flexibility, and adhesion. In fact, adhesion of a cell to its surrounding matrix is required for cell survival except in the case of metastatic tumor cells, o which have overcome the need for cell-ECM anchorage. This phenomenon suggests that the ECM plays a critical role in the molecular mechanisms of growth control and metastasis. (Reviewed in Ruoslahti, E. (1996) Sci. An. 275:72-77.) Furthermore, the ECM determines the structure and physical properties of connective tissue and is particularly important for moφhogenesis and other processes associated with embryonic development and pattern formation. The collagens comprise a family of ECM proteins that provide structure to bone, teeth, skin, Ugaments, tendons, cartilage, blood vessels, and basement membranes. Multiple collagen proteins have been identified. Three collagen molecules fold together in a triple helix stabihzed by interchain disulfide bonds. Bundles of these triple heUces then associate to form fibrils. Collagen primary structure 5 consists of hundreds of (Gly-X-Y) repeats where about a third of the X and Y residues are Pro.
Glycines are crucial to helix formation as the bulkier amino acid sidechains cannot fold into the triple helical conformation. Because of these strict sequence requirements, mutations in collagen genes have severe consequences. Osteogenesis imperfecta patients have brittle bones that fracture easily; in severe cases patients die in utero or at birth. Ehlers-Danlos syndrome patients have hyperelastic skin, 0 hypermobile joints, and susceptibiUty to aortic and intestinal rupture. Chondrodysplasia patients have short stature and ocular disorders. Aport syndrome patients have hematuria, sensorineural deafness, and eye lens deformation. (Isselbacher, KJ. et al. (1994) Harrison's Principles of Internal Medicine, McGraw-Hill, Inc., New York NY, pp. 2105-2117; and Creighton, T.E. (1984) Proteins, Structures and Molecular Principles. W.H. Freeman and Company, New York NY, pp. 191-197.) 5 Elastin and related proteins confer elasticity to tissues such as skin, blood vessels, and lungs.
Elastin is a highly hydrophobic protein of about 750 amino acids that is rich in proUne and glycine residues. Elastin molecules are highly cross-Unked, forming an extensive extracellular network of fibers and sheets. Elastin fibers are surrounded by a sheath of microfibrils which are composed of a number of glycoproteins, including fibrilUn. Mutations in the gene encoding fibrilUn are responsible for o Marfan' s syndrome, a genetic disorder characterized by defects in connective tissue. In severe cases, the aortas of afflicted individuals are prone to rupture. (Reviewed in Aberts, supra, pp. 984-986.)
Fibronectin is a large ECM glycoprotein found in all vertebrates. Fibronectin exists as a dimer of two subunits, each containing about 2,500 amino acids. Each subunit folds into a rod-Uke structure containing multiple domains. The domains each contain multiple repeated modules, the most common 5 of which is the type III fibronectin repeat. The type III fibronectin repeat is about 90 amino acids in length and is also found in other ECM proteins and in some plasma membrane and cytoplasmic proteins. Furthermore, some type III fibronectin repeats contain a characteristic tripeptide consisting of Aginine-Glycine-Aspartic acid (RGD). The RGD sequence is recognized by the integrin family of cell surface receptors and is also found in other ECM proteins. Disruption of both copies of the gene o encoding fibronectin causes early embryonic lethaUty in mice. The mutant embryos display extensive moφhological defects, including defects in the formation of the notochord, somites, heart, blood vessels, neural tube, and extraembryonic structures. (Reviewed in Aberts, supra, pp. 986-987.)
Laminin is a major glycoprotein component of the basal lamina which underlies and supports epitheUal cell sheets. Laminin is one of the first ECM proteins synthesized in the developing embryo. Laminin is an 850 kilodalton protein composed of three polypeptide chains joined in the shape of a cross by disulfide bonds. Laminin is especially important for angiogenesis and in particular, for guiding the formation of capillaries. (Reviewed in Aberts, supra, pp. 990-991.)
There are many other types of proteinaceous ECM components, most of which can be 5 classified as proteoglycans. Proteoglycans are composed of unbranched polysaccharide chains
(glycosaminoglycans) attached to protein cores. Common proteoglycans include aggrecan, betaglycan, decorin, perlecan, serglycin, and syndecan-1. Some of these molecules not only provide mechanical support, but also bind to extracellular signaUng molecules, such as fibroblast growth factor and transforming growth factor β, suggesting a role for proteoglycans in cell-cell communication and cell 0 growth. (Reviewed in Aberts, supra, pp. 973-978.) Likewise, the glycoproteins tenascin-C and tenascin-R are expressed in developing and lesioned neural tissue and provide stimulatory and anti- adhesive (inhibitory) properties, respectively, for axonal growth. (Faissner, A. (1997) Cell Tissue Res. 290:331-341.)
5 Cytoskeletal Molecules
SEQ ID NO:40, SEQ ID NO:41, SEQ ID NO:42, SEQ ID NO:43, SEQ ID NO:44, and SEQ ID NO:45 encode, for example, cytoskeletal molecules.
The cytoskeleton is a cytoplasmic network of protein fibers that mediate cell shape, structure, and movement. The cytoskeleton supports the cell membrane and forms tracks along which o organelles and other elements move in the cytosol. The cytoskeleton is a dynamic structure that allows cells to adopt various shapes and to carry out directed movements. Major cytoskeletal fibers include the microtubules, the microfilaments, and the intermediate filaments. Motor proteins, including myosin, dynein, and kinesin, drive movement of or along the fibers. The motor protein dynamin drives the formation of membrane vesicles. Accessory or associated proteins modify the 5 structure or activity of the fibers while cytoskeletal membrane anchors connect the fibers to the cell membrane. Tubulins
Microtubules, cytoskeletal fibers with a diameter of about 24 nm, have multiple roles in the cell. Bundles of microtubules form cilia and flagella, which are whip-like extensions of the cell o membrane that are necessary for sweeping materials across an epithelium and for swimming of sperm, respectively. Marginal bands of microtubules in red blood cells and platelets are important for these cells' pliability. Organelles, membrane vesicles, and proteins are transported in the cell along tracks of microtubules. For example, microtubules run through nerve cell axons, allowing bidirectional transport of materials and membrane vesicles between the cell body and the nerve terminal. Failure to supply the nerve terminal with these vesicles blocks the transmission of neural signals. Microtubules are also critical to chromosomal movement during cell division. Both stable and short-lived populations of microtubules exist in the cell.
Microtubules are polymers of GTP-binding tubulin protein subunits. Each subunit is a 5 heterodimer of α- and β- tubulin, multiple isoforms of which exist. The hydrolysis of GTP is linked to the addition of tubuUn subunits at the end of a microtubule. The subunits interact head to tail to form protofilaments; the protofilaments interact side to side to form a microtubule. A microtubule is polarized, one end ringed with α-tubulin and the other with β-tubulin, and the two ends differ in their rates of assembly. Generally, each microtubule is composed of 13 protofilaments although 11 or 15 0 protofilament-microtubules are sometimes found. Cilia and flagella contain doublet microtubules. Microtubules grow from speciaUzed structures known as centrosomes or microtubule-organizing centers (MTOCs). MTOCs may contain one or two centrioles, which are pinwheel arrays of triplet microtubules. The basal body, the organizing center located at the base of a cilium or flagellum, contains one centriole. Gamma tubulin present in the MTOC is important for nucleating the 5 polymerization of α- and β- tubulin heterodimers but does not polymerize into microtubules. Microtubule- Associated Proteins
Microtubule-associated proteins (MAPs) have roles in the assembly and stabilization of microtubules. One major family of MAPs, assembly MAPs, can be identified in neurons as well as non-neuronal cells. Assembly MAPs are responsible for cross-Unking microtubules in the cytosol. o These MAPs are organized into two domains: a basic microtubule-binding domain and an acidic projection domain. The projection domain is the binding site for membranes, intermediate filaments, or other microtubules. Based on sequence analysis, assembly MAPs can be further grouped into two types: Type I and Type II. Type I MAPs, which include MAPI A and MAP1B, are large, filamentous molecules that co-purify with microtubules and are abundantly expressed in brain and testes. Type I 5 MAPs contain several repeats of a positively-charged amino acid sequence motif that binds and neutralizes negatively charged tubulin, leading to stabiUzation of microtubules. MAPI A and MAP1B are each derived from a single precursor polypeptide that is subsequently proteolytically processed to generate one heavy chain and one Ught chain.
Another Ught chain, LC 3, is a 16.4 kDa molecule that binds MAPI A, MAP1B, and o microtubules. It is suggested that LC3 is synthesized from a source other than the MAPI A or MAPI B transcripts, and that the expression of LC3 may be important in regulating the microtubule binding activity of MAPI A and MAP1B during cell proUferation (Mann, S.S. et al. (1994) J. Biol. Chem. 269:11492-11497).
Type II MAPs, which include MAP2a, MAP2b, MAP2c, MAP4, and Tau, are characterized by three to four copies of an 18-residue sequence in the microtubule-binding domain. MAP2a, MAP2b, and MAP2c are found only in dendrites, MAP4 is found in non-neuronal cells, and Tau is found in axons and dendrites of nerve cells. Ater native spUcing of the Tau mRNA leads to the existence of multiple forms of Tau protein. Tau phosphorylation is altered in neurodegenerative disorders such as Azheimer' s disease, Pick's disease, progressive supranuclear palsy, corticobasal degeneration, and familial frontotemporal dementia and Parkinsonism linked to chromosome 17. The altered Tau phosphorylation leads to a collapse of the microtubule network and the formation of intraneuronal Tau aggregates (Spillantini, M.G. and M. Goedert (1998) Trends Neurosci. 21:428-433).
The protein pericentrin is found in the MTOC and has a role in microtubule assembly. Actins
Microfilaments, cytoskeletal filaments with a diameter of about 7-9 nm, are vital to cell locomotion, cell shape, cell adhesion, cell division, and muscle contraction. Assembly and disassembly of the microfilaments allow cells to change their moφhology. Microfilaments are the polymerized form of actin, the most abundant intracellular protein in the eukaryotic cell. Human cells contain six isoforms of actin. The three α-actins are found in different kinds of muscle, nonmuscle β- actin and nonmuscle γ-actin are found in nonmuscle cells, and another γ-actin is found in intestinal smooth muscle cells. G-actin, the monomeric form of actin, polymerizes into polarized, helical F- actin filaments, accompanied by the hydrolysis of ATP to ADP. Actin filaments associate to form bundles and networks, providing a framework to support the plasma membrane and determine cell shape. These bundles and networks are connected to the cell membrane. In muscle cells, thin filaments containing actin slide past thick filaments containing the motor protein myosin during contraction. A family of actin-related proteins exist that are not part of the actin cytoskeleton, but rather associate with microtubules and dynein. Actin- Associated Proteins Actin-associated proteins have roles in cross-linking, severing, and stabilization of actin filaments and in sequestering actin monomers. Several of the actin-associated proteins have multiple functions. Bundles and networks of actin filaments are held together by actin cross-linking proteins. These proteins have two actin-binding sites, one for each filament. Short cross-linking proteins promote bundle formation while longer, more flexible cross-linking proteins promote network formation. Calmodulin-like calcium-binding domains in actin cross-linking proteins allow calcium regulation of cross-linking. Group I cross-linking proteins have unique actin-binding domains and include the 30 kD protein, EF-la, fascin, and scruin. Group II cross-linking proteins have a 7,000- MW actin-binding domain and include villin and dematin. Group III cross-Unking proteins have pairs of a 26,000-MW actin-binding domain and include fimbrin, spectrin, dystrophin, ABP 120, and filamin. Severing proteins regulate the length of actin filaments by breaking them into short pieces or by blocking their ends. Severing proteins include gCAP39, severin (fragmin), gelsolin, and vilUn. Capping proteins can cap the ends of actin filaments, but cannot break filaments. Capping proteins include CapZ and tropomodulin. The proteins thymosin and profilin sequester actin monomers in the cytosol, allowing a pool of unpolymerized actin to exist. The actin-associated proteins tropomyosin, troponin, and caldesmon regulate muscle contraction in response to calcium. Intermediate Filaments and Associated Proteins
Intermediate filaments (IFs) are cytoskeletal fibers with a diameter of about 10 nm, intermediate between that of microfilaments and microtubules. IFs serve structural roles in the cell, reinforcing cells and organizing cells into tissues. IFs are particularly abundant in epidermal cells and in neurons. IFs are extremely stable, and, in contrast to microfilaments and microtubules, do not function in cell motility.
Five types of IF proteins are known in mammals. Type I and Type II proteins are the acidic and basic keratins, respectively. Heterodimer s of the acidic and basic keratins are the building blocks of keratin IFs. Keratins are abundant in soft epitheUa such as skin and cornea, hard epitheUa such as nails and hair, and in epitheUa that Une internal body cavities. Mutations in keratin genes lead to epitheUal diseases including epidermolysis bullosa simplex, bullous congenital ichthyosiform erythroderma (epidermolytic hyperkeratosis), non-epidermolytic and epidermolytic palmoplantar keratoderma, ichthyosis bullosa of Siemens, pachyonychia congenita, and white sponge nevus. Some of these diseases result in severe skin bUstering. (See, e.g., Wawersik, M. et al. (1997) J. Biol. Chem. 272:32557-32565; and Corden L.D. and W.H. McLean (1996) Exp. Dermatol. 5:297-307.)
Type III IF proteins include desmin, gUal fibrillary acidic protein, vimentin, and peripherin. Desmin filaments in muscle cells link myofibrils into bundles and stabiUze sarcomeres in contracting muscle. GUal fibrillary acidic protein filaments are found in the gUal cells that surround neurons and astrocytes. Vimentin filaments are found in blood vessel endothelial cells, some epitheUal cells, and mesenchymal cells such as fibroblasts, and are commonly associated with microtubules. Vimentin filaments may have roles in keeping the nucleus and other organelles in place in the cell. Type IV IFs include the neurofilaments and nestin. Neurofilaments, composed of three polypeptides NF-L, NF-M, and NF-H, are frequently associated with microtubules in axons. Neurofilaments are responsible for the radial growth and diameter of an axon, and ultimately for the speed of nerve impulse transmission. Changes in phosphorylation and metabolism of neurofilaments are observed in neurodegenerative diseases including amyotrophic lateral sclerosis, Parkinson's disease, and Azheimer 's disease (Julien, J.P. and W.E. Mushynski (1998) Prog. Nucleic Acid Res. Mol. Biol. 61:1-23). Type V IFs, the lamins, are found in the nucleus where they support the nuclear membrane. IFs have a central α-heUcal rod region interrupted by short nonheUcal Unker segments. The rod region is bracketed, in most cases, by non-heUcal head and tail domains. The rod regions of intermediate filament proteins associate to form a coiled-coil dimer. A highly ordered assembly process leads from the dimers to the IFs. Neither ATP nor GTP is needed for IF assembly, unUke that of 5 microfilaments and microtubules.
IF-associated proteins (IFAPs) mediate the interactions of IFs with one another and with other cell structures. IFAPs cross-link IFs into a bundle, into a network, or to the plasma membrane, and may cross-Unk IFs to the microfilament and microtubule cytoskeleton. Microtubules and IFs are in particular closely associated. IFAPs include BPAG1, plakoglobin, desmoplakin I, desmoplakin II, o plectin, ankyrin, filaggrin, and lamin B receptor.
Cvtoskeletal-Membrane Anchors
Cytoskeletal fibers are attached to the plasma membrane by specific proteins. These attachments are important for maintaining cell shape and for muscle contraction. In erythrocytes, the spectrin- actin cytoskeleton is attached to cell membrane by three proteins, band 4.1, ankyrin, and 5 adducin. Defects in this attachment result in abnormally shaped cells which are more rapidly degraded by the spleen, leading to anemia. In platelets, the spectrin-actin cytoskeleton is also linked to the membrane by ankyrin; a second actin network is anchored to the membrane by filamin. In muscle cells the protein dystrophin links actin filaments to the plasma membrane; mutations in the dystrophin gene lead to Duchenne muscular dystrophy. In adherens junctions and adhesion plaques o the peripheral membrane proteins α-actinin and vinculin attach actin filaments to the cell membrane.
IFs are also attached to membranes by cytoskeletal-membrane anchors. The nuclear lamina is attached to the inner surface of the nuclear membrane by the lamin B receptor. Vimentin IFs are attached to the plasma membrane by ankyrin and plectin. Desmosome and hemidesmosome membrane junctions hold together epithelial cells of organs and skin. These membrane junctions 5 allow shear forces to be distributed across the entire epithelial cell layer, thus providing strength and rigidity to the epithelium. IFs in epithelial cells are attached to the desmosome by plakoglobin and desmoplakins. The proteins that Unk IFs to hemidesmosomes are not known. Desmin IFs surround the sarcomere in muscle and are linked to the plasma membrane by paranemin, synemin, and ankyrin. Myosin-related Motor Proteins o Myosins are actin-activated ATPases, found in eukaryotic cells, that couple hydrolysis of
ATP with motion. Myosin provides the motor function for muscle contraction and intracellular movements such as phagocytosis and rearrangement of cell contents during mitotic cell division (cytokinesis). The contractile unit of skeletal muscle, termed the sarcomere, consists of highly ordered arrays of thin actin-containing filaments and thick myosin-containing filaments. Crossbridges form between the thick and thin filaments, and the ATP-dependent movement of myosin heads within the thick filaments pulls the thin filaments, shortening the sarcomere and thus the muscle fiber.
Myosins are composed of one or two heavy chains and associated light chains. Myosin heavy chains contain an amino-terminal motor or head domain, a neck that is the site of light-chain 5 binding, and a carboxy-terminal tail domain. The tail domains may associate to form an α-helical coiled coil. Conventional myosins, such as those found in muscle tissue, are composed of two myosin heavy-chain subunits, each associated with two light-chain subunits that bind at the neck region and play a regulatory role. Unconventional myosins, believed to function in intracellular motion, may contain either one or two heavy chains and associated light chains. There is evidence for o about 25 myosin heavy chain genes in vertebrates, more than half of them unconventional.
Dynein-related Motor Proteins
Dyneins are (-) end-directed motor proteins which act on microtubules. Two classes of dyneins, cytosoUc and axonemal, have been identified. CytosoUc dyneins are responsible for translocation of materials along cytoplasmic microtubules, for example, transport from the nerve 5 terminal to the cell body and transport of endocytic vesicles to lysosomes. Cytoplasmic dyneins are also reported to play a role in mitosis. Axonemal dyneins are responsible for the beating of flagella and ciUa. Dynein on one microtubule doublet walks along the adjacent microtubule doublet. This sUding force produces bending forces that cause the flagellum or ciUum to beat. Dyneins have a native mass between 1000 and 2000 kDa and contain either two or three force-producing heads driven by the o hydrolysis of ATP. The heads are Unked via stalks to a basal domain which is composed of a highly variable number of accessory intermediate and Ught chains.
Kinesin-related Motor Proteins
Kinesins are (+) end-directed motor proteins which act on microtubules. The prototypical kinesin molecule is involved in the transport of membrane-bound vesicles and organelles. This function 5 is particularly important for axonal transport in neurons. Kinesin is also important in all cell types for the transport of vesicles from the Golgi complex to the endoplasmic reticulum. This role is critical for maintaining the identity and functionaUty of these secretory organelles.
Kinesins define a ubiquitous, conserved family of over 50 proteins that can be classified into at least 8 subfamiUes based on primary amino acid sequence, domain structure, velocity of movement, and 0 cellular function. (Reviewed in Moore, J.D. and S.A. Endow (1996) Bioessays 18:207-219; and Hoyt,
AM. (1994) Curr. Opia Cell Biol. 6:63-68.) The prototypical kinesin molecule is a heterotetramer comprised of two heavy polypeptide chains (KHCs) and two light polypeptide chains (KLCs). The
KHC subunits are typically referred to as "kinesin." KHC is about 1000 amino acids in length, and
KLC is about 550 amino acids in length. Two KHCs dimerize to form a rod-shaped molecule with three distinct regions of secondary structure. At one end of the molecule is a globular motor domain that functions in ATP hydrolysis and microtubule binding. Kinesin motor domains are highly conserved and share over 70% identity. Beyond the motor domain is an α-hehcal coiled-coil region which mediates dimerization. At the other end of the molecule is a fan-shaped tail that associates with 5 molecular cargo. The tail is formed by the interaction of the KHC C-termini with the two KLCs. Members of the more divergent subfamiUes of kinesins are called kinesin-related proteins (KRPs), many of which function during mitosis in eukaryotes (Hoyt, supra). Some KRPs are required for assembly of the mitotic spindle. In vivo and in vitro analyses suggest that these KRPs exert force on microtubules that comprise the mitotic spindle, resulting in the separation of spindle poles. o Phosphorylation of KRP is required for this activity. Failure to assemble the mitotic spindle results in abortive mitosis and chromosomal aneuploidy, the latter condition being characteristic of cancer cells. In addition, a unique KRP, centromere protein E, localizes to the kinetochore of human mitotic chromosomes and may play a role in their segregation to opposite spindle poles. Dynamin-related Motor Proteins 5 Dynamin is a large GTPase motor protein that functions as a "molecular pinchase," generating a mechanochemical force used to sever membranes. This activity is important in forming clathrin-coated vesicles from coated pits in endocytosis and in the biogenesis of synaptic vesicles in neurons. Binding of dynamin to a membrane leads to dynamin' s self-assembly into spirals that may act to constrict a flat membrane surface into a tubule. GTP hydrolysis induces a change in o conformation of the dynamin polymer that pinches the membrane tubule, leading to severing of the membrane tubule and formation of a membrane vesicle. Release of GDP and inorganic phosphate leads to dynamin disassembly. Following disassembly the dynamin may either dissociate from the membrane or remain associated to the vesicle and be transported to another region of the cell. Three homologous dynamin genes have been discovered, in addition to several dynamin-related proteins. 5 Conserved dynamin regions are the N-terminal GTP-binding domain, a central pleckstrin homology domain that binds membranes, a central coiled-coil region that may activate dynamin' s GTPase activity, and a C-terminal proline-rich domain that contains several motifs that bind SH3 domains on other proteins. Some dynamin-related proteins do not contain the pleckstrin homology domain or the proline-rich domain. (See McNiven, M.A. (1998) Cell 94:151-154; Scaife, R.M. and R.L. Margolis 0 (1997) Cell. Signal. 9:395-401.)
The cytoskeleton is reviewed in Lodish, H. et al. (1995) Molecular Cell Biology, Scientific American Books, New York NY.
Ribosomal Molecules SEQ ID NO:49, SEQ ID NO:50, SEQ ID NO:51, SEQ ID NO:52, and SEQ ID NO:53 encode, for example, ribosomal molecules.
Ribosomal RNAs (rRNAs) are assembled, along with ribosomal proteins, into ribosomes, which are cytoplasmic particles that translate messenger RNA into polypeptides. The eukaryotic ribosome is composed of a 60S (large) subunit and a 40S (small) subunit, which together form the 80S ribosome. In addition to the 18S, 28S, 5S, and 5.8S rRNAs, the ribosome also contains more than fifty proteins. The ribosomal proteins have a prefix which denotes the subunit to which they belong, either L (large) or S (small). Ribosomal protein activities include binding rRNA and organizing the conformation of the junctions between rRNA helices (Woodson, S.A. and N.B. Leontis (1998) Curr. Opin. Struct. Biol. 8:294-300; Ramakrishnan, V. and S.W. White (1998) Trends Biochem. Sci. 23:208-212.) Three important sites are identified on the ribosome. The aminoacyl- tRNA site (A site) is where charged tRNAs (with the exception of the initiator-tRNA) bind on arrival at the ribosome. The peptidyl-tRNA site (P site) is where new peptide bonds are formed, as well as where the initiator tRNA binds. The exit site (E site) is where deacylated tRNAs bind prior to their release from the ribosome. (The ribosome is reviewed in Stryer, L. (1995) Biochemistry W.H. Freeman and Company, New York NY, pp. 888-908; and Lodish, H. et al. (1995) Molecular Cell Biology Scientific American Books, New York NY. pp. 119-138.)
Chromatin Molecules The nuclear DNA of eukaryotes is organized into chromatin. Two types of chromatin are observed: euchromatin, some of which may be transcribed, and heterochromatin so densely packed that much of it is inaccessible to transcription. Chromatin packing thus serves to regulate protein expression in eukaryotes. Bacteria lack chromatin and the chromatin-packing level of gene regulation. The fundamental unit of chromatin is the nucleosome of 200 DNA base pairs associated with two copies each of histones H2A, H2B, H3, and H4. Adjascent nucleosomes are Unked by another class of histones, HI . Low molecular weight non-histone proteins called the high mobiUty group (HMG), associated with chromatin, may function in the unwinding of DNA and stabiUzation of single- stranded DNA Chromodomain proteins function in compaction of chromatin into its transcriptionally silent heterochromatin form. During mitosis, all DNA is compacted into heterochromatin and transcription ceases.
Transcription in inte hase begins with the activation of a region of chromatin. Active chromatin is decondensed. Decondensation appears to be accompanied by changes in binding coefficient, phosphorylation and acetylation states of chromatin histones. HMG proteins HMG13 and HMG17 selectively bind activated chromatin. Topoisomerases remove superhelical tension on DNA The activated region decondenses, allowing gene regulatory proteins and transcription factors to assemble on the DNA.
Patterns of chromatin structure can be stably inherited, producing heritable patterns of gene expression. In mammals, one of the two X chromosomes in each female cell is inactivated by 5 condensation to heterochromatin during zygote development. The inactive state of this chromosome is inherited, so that adult females are mosaics of clusters of paternal-X and maternal-X clonal cell groups. The condensed X chromosome is reactivated in meiosis.
Chromatin is associated with disorders of protein expression such as thalassemia, a genetic anemia resulting from the removal of the locus control region (LCR) required for decondensation of the o globin gene locus.
For a review of chromatin structure and function see Aberts, B. et al. (1994) Molecular Cell Biology, third edition, Garland Publishing, Inc., New York NY, pp. 351-354, 433-439.
Electron Transfer Associated Molecules 5 Electron carriers such as cytochromes accept electrons from NADH or FADHj and donate them to other electron carriers. Most electron-transferring proteins, except ubiquinone, are prosthetic groups such as flavins, heme, FeS clusters, and copper, bound to inner membrane proteins. Adrenodoxin, for example, is an FeS protein that forms a complex with NADPH :adrenodoxin reductase and cytochrome p450. Cytochromes contain a heme prosthetic group, a poφhyrin ring o containing a tightly bound iron atom. Electron transfer reactions play a crucial role in cellular energy production.
Energy is produced by the oxidation of glucose and fatty acids. Glucose is initially converted to pyruvate in the cytoplasm. Fatty acids and pyruvate are transported to the mitochondria for complete oxidation to C02 coupled by enzymes to the transport of electrons from NADH and FADH2 5 to oxygen and to the synthesis of ATP (oxidative phosphorylation) from ADP and P,.
Pyruvate is transported into the mitochondria and converted to acetyl-CoA for oxidation via the citric acid cycle, involving pyruvate dehydrogenase components, dihydrolipoyl fransacetylase, and dihydrolipoyl dehydrogenase. Enzymes involved in the citric acid cycle include: citrate synthetase, aconitases, isocitrate dehydrogenase, alpha-ketoglutarate dehydrogenase complex including o transsuccinylases, succinyl CoA synthetase, succinate dehydrogenase, fumarases, and malate dehydrogenase. Acetyl CoA is oxidized to C02 with concomitant formation of NADH, FADH2, and GTP. In oxidative phosphorylation, the transfer of electrons from NADH and FADH2 to oxygen by dehydrogenases is coupled to the synthesis of ATP from ADP and P, by the FQF, ATPase complex in the mitochondrial inner membrane. Enzyme complexes responsible for electron transport and ATP synthesis include the FQFJ ATPase complex, ubiquinone(CoQ)-cytochrome c reductase, ubiquinone reductase, cytochrome b, cytochrome c1; FeS protein, and cytochrome c oxidase.
ATP synthesis requires membrane transport enzymes including the phosphate transporter and the ATP- ADP antiport protein. The ATP-binding casette (ABC) superfamily has also been suggested 5 as belonging to the mitochondrial transport group (Hogue, D.L. et al. (1999) J. Mol. Biol. 285:379- 389). Brown fat uncoupling protein dissipates oxidative energy as heat, and may be involved the fever response to infection and trauma (Cannon, B. et al. (1998) Ann. NY Acad. Sci. 856:171-187).
Mitochondria are oval-shaped organelles comprising an outer membrane, a tightly folded inner membrane, an intermembrane space between the outer and inner membranes, and a matrix o inside the inner membrane. The outer membrane contains many porin molecules that allow ions and charged molecules to enter the intermembrane space, while the inner membrane contains a variety of transport proteins that transfer only selected molecules. Mitochondria are the primary sites of energy production in cells.
Mitochondria contain a small amount of DNA. Human mitochondrial DNA encodes 13 5 proteins, 22 tRNAs, and 2 rRNAs. Mitochondrial-DNA encoded proteins include NADH-Q reductase, a cytochrome reductase subunit, cytochrome oxidase subunits, and ATP synthase subunits.
Electron-transfer reactions also occur outside the mitochondria in locations such as the endoplasmic reticulum, which plays a crucial role in Upid and protein biosynthesis. Cytochrome b5 is a central electron donor for various reductive reactions occurring on the cytoplasmic surface of liver o endoplasmic reticulum. Cytochrome b5 has been found in Golgi, plasma, endoplasmic reticulum
(ER), and microbody membranes.
For a review of mitochondrial metabolism and regulation, see Lodish, H. et al. (1995) Molecular Cell Biology. Scientific American Books, New York NY, pp. 745-797 and Stryer (1995) Biochemistry. W.H. Freeman and Co., San Francisco CA, PP 529-558, 988-989. 5 The majority of mitochondrial proteins are encoded by nuclear genes, are synthesized on cytosoUc ribosomes, and are imported into the mitochondria. Nuclear-encoded proteins which are destined for the mitochondrial matrix typically contain positively-charged amino terminal signal sequences. Import of these preproteins from the cytoplasm requires a multisubunit protein complex in the outer membrane known as the translocase of outer mitochondrial membrane (TOM; previously o designated MOM; Pfanner, N. et al. (1996) Trends Biochem. Sci. 21 :51-52) and at least three inner membrane proteins which comprise the translocase of inner mitochondrial membrane (TIM; previously designated MIM; Pfanner, supra). An inside-negative membrane potential across the inner mitochondrial membrane is also required for preprotein import. Preproteins are recognized by surface receptor components of the TOM complex and are translocated through a proteinaceous pore formed 5 by other TOM components. Proteins targeted to the matrix are then recognized by the import machinery of the TIM complex. The import systems of the outer and inner membranes can function independently (Segui-Real, B. et al. (1993) EMBO J. 12:2211-2218).
Once precursor proteins are in the mitochondria, the leader peptide is cleaved by a signal peptidase to generate the mature protein. Most leader peptides are removed in a one step process by a protease termed mitochondrial processing peptidase (MPP) (Paces, V. et al. (1993) Proc. Natl. Acad. Sci. USA 90:5355-5358). In some cases a two-step process occurs in which MPP generates an intermediate precursor form which is cleaved by a second enzyme, mitochondrial intermediate peptidase, to generate the mature protein.
Mitochondrial dysfunction leads to impaired calcium buffering, generation of free radicals that may participate in deleterious intracellular and extracellular processes, changes in mitochondrial permeability and oxidative damage which is observed in several neurodegenerative diseases. Neurodegenerative diseases Unked to mitochondrial dysfunction include some forms of Alzheimer's disease, Friedreich's ataxia, familial amyotrophic lateral sclerosis, and Huntington's disease (Beal, M.F. (1998) Biochim. Biophys. Acta 1366:211-213). The myocardium is heavily dependent on oxidative metabolism, so mitochondrial dysfunction often leads to heart disease (DiMauro, S. and M. Hirano (1998) Curr. Opin. Cardiol 13:190-197). Mitochondria are implicated in disorders of cell proliferation, since they play an important role in a cell's decision to proliferate or self-destruct through apoptosis. The oncoprotein Bcl-2, for example, promotes cell proliferation by stabilizing mitochondrial membranes so that apoptosis signals are not released (Susin, S.A. (1998) Biochim. Biophys. Acta 1366:151-165).
Transcription Factor Molecules
SEQ ID NO:19, SEQ ID NO:20, SEQ ID NO:21, SEQ ID NO:22, SEQ ID NO:23, SEQ ID NO:24, SEQ ID NO:25, SEQ ID NO:26, SEQ ID NO:27, SEQ ID NO:28, SEQ ID NO:29, SEQ ID NO:30, SEQ ID NO:31, SEQ ID NO:32, and SEQ ID NO:33 encode, for example, transcription factor molecules.
Multicellular organisms are comprised of diverse cell types that differ dramatically both in structure and function. The identity of a cell is determined by its characteristic pattern of gene expression, and different cell types express overlapping but distinctive sets of genes throughout development. Spatial and temporal regulation of gene expression is critical for the control of cell proliferation, cell differentiation, apoptosis, and other processes that contribute to organismal development. Furthermore, gene expression is regulated in response to extracellular signals that mediate cell-cell communication and coordinate the activities of different cell types. Appropriate gene regulation also ensures that cells function efficiently by expressing only those genes whose functions are required at a given time. Transcriptional regulatory proteins are essential for the control of gene expression. Some of these proteins function as transcription factors that initiate, activate, repress, or terminate gene transcription. Transcription factors generally bind to the promoter, enhancer, and upstream regulatory regions of a gene in a sequence-specific manner, although some factors bind regulatory elements within or downstream of a gene's coding region. Transcription factors may bind to a specific region of DNA singly or as a complex with other accessory factors. (Reviewed in Lewin, B. (1990) Genes IV, Oxford University Press, New York NY, and Cell Press, Cambridge MA, pp. 554-570.)
The double helix structure and repeated sequences of DNA create topological and chemical features which can be recognized by transcription factors. These features are hydrogen bond donor and acceptor groups, hydrophobic patches, major and minor grooves, and regular, repeated stretches of sequence which induce distinct bends in the helix. Typically, transcription factors recognize specific DNA sequence motifs of about 20 nucleotides in length. Multiple, adjacent transcription factor-binding motifs may be required for gene regulation.
Many transcription factors incoφorate DNA-binding structural motifs which comprise either α heUces or β sheets that bind to the major groove of DNA. Four well-characterized structural motifs are helix-turn-helix, zinc finger, leucine zipper, and helix-loop-helix. Proteins containing these motifs may act alone as monomers, or they may form homo- or heterodimers that interact with DNA.
The helix-turn-helix motif consists of two α helices connected at a fixed angle by a short chain of amino acids. One of the helices binds to the major groove. Helix-turn-helix motifs are exemplified by the homeobox motif which is present in homeodomain proteins. These proteins are critical for specifying the anterior-posterior body axis during development and are conserved throughout the animal kingdom. The Antennapedia and Ultrabithorax proteins of Drosophila melanogaster are prototypical homeodomain proteins (Pabo, CO. and R.T. Sauer (1992) Annu. Rev. Biochem. 61:1053-1095). The zinc finger motif, which binds zinc ions, generally contains tandem repeats of about 30 amino acids consisting of periodically spaced cysteine and histidine residues. Examples of this sequence pattern, designated C2H2 and C3HC4 ("RING" finger), have been described (Lewin, supra). Zinc finger proteins each contain an α heUx and an antiparallel β sheet whose proximity and conformation are maintained by the zinc ion. Contact with DNA is made by the arginine prece ding the α heUx and by the second, third, and sixth residues of the α heUx. Variants of the zinc finger motif include poorly defined cysteine-rich motifs which bind zinc or other metal ions. These motifs may not contain histidine residues and are generally nonrepetitive.
The leucine zipper motif comprises a stretch of amino acids rich in leucine which can form an amphipafhic α helix. This structure provides the basis for dimerization of two leucine zipper proteins. The region adjacent to the leucine zipper is usually basic, and upon protein dimerization, is optimally positioned for binding to the major groove. Proteins containing such motifs are generally referred to as bZIP transcription factors.
The helix-loop-heUx motif (HLH) consists of a short α heUx connected by a loop to a longer 5 cc heUx. The loop is flexible and allows the two helices to fold back against each other and to bind to DNA. The transcription factor Myc contains a prototypical HLH motif.
Most transcription factors contain characteristic DNA binding motifs, and variations on the above motifs and new motifs have been and are currently being characterized (Faisst, S. and S. Meyer (1992) Nucleic Acids Res. 20:3-26). 0 Many neoplastic disorders in humans can be attributed to inappropriate gene expression.
Malignant cell growth may result from either excessive expression of tumor promoting genes or insufficient expression of tumor suppressor genes (Cleary, M.L. (1992) Cancer Surv. 15:89-104). Chromosomal translocations may also produce chimeric loci which fuse the coding sequence of one gene with the regulatory regions of a second unrelated gene. Such an arrangement Ukely results in 5 inappropriate gene transcription, potentially contributing to malignancy.
In addition, the immune system responds to infection or trauma by activating a cascade of events that coordinate the progressive selection, amplification, and mobilization of cellular defense mechanisms. A complex and balanced program of gene activation and repression is involved in this process. However, hyperactivity of the immune system as a result of improper or insufficient o regulation of gene expression may result in considerable tissue or organ damage. This damage is well documented in immunological responses associated with arthritis, allergens, heart attack, stroke, and infections (Isselbacher, KJ. et al. (1996) Harrison's Principles of Internal Medicine, 13/e, McGraw Hill, Inc. and Teton Data Systems Software).
Furthermore, the generation of multicellular organisms is based upon the induction and 5 coordination of cell differentiation at the appropriate stages of development. Central to this process is differential gene expression, which confers the distinct identities of cells and tissues throughout the body. Failure to regulate gene expression during development can result in developmental disorders. Human developmental disorders caused by mutations in zinc finger-type transcriptional regulators include: urogenenital developmental abnormalities associated with WT1 ; Greig o cephalopolysyndactyly, Pallister-Hall syndrome, and postaxial polydactyly type A (GLI3); and
Townes-Brocks syndrome, characterized by anal, renal, limb, and ear abnormalities (SALL1) (Engelkamp, D. and V. van Heyningen (1996) Curr. Opin. Genet. Dev. 6:334-342; Kohlhase, J. et al. (1999) Am. J. Hum. Genet. 64:435-445). Cell Membrane Molecules
SEQ ID NO:46, SEQ ID NO:47, and SEQ ID NO:48 encode, for example, cell membrane molecules.
Eukaryotic cells are surrounded by plasma membranes which enclose the cell and maintain an 5 environment inside the cell that is distinct from its surroundings. In addition, eukaryotic organisms are distinct from prokaryotes in possessing many intracellular organelle and vesicle structures. Many of the metabolic reactions which distinguish eukaryotic biochemistry from prokaryotic biochemistry take place within these structures. The plasma membrane and the membranes surrounding organeUes and vesicles are composed of phosphoglycerides, fatty acids, cholesterol, phosphoUpids, glycolipids, 0 proteoglycans, and proteins. These components confer identity and functionaUty to the membranes with which they associate. Integral Membrane Proteins
The majority of known integral membrane proteins are transmembrane proteins (TM) which are characterized by an extracellular, a transmembrane, and an intracellular domain. TM domains are 5 typically comprised of 15 to 25 hydrophobic amino acids which are predicted to adopt an α-helical conformation. TM proteins are classified as bitopic (Types I and II) and polytopic (Types III and IV) (Singer, S.J. (1990) Annu. Rev. Cell Biol. 6:247-296). Bitopic proteins span the membrane once while polytopic proteins contain multiple membrane-spanning segments. TM proteins function as cell-surface receptors, receptor-interacting proteins, transporters of ions or metabolites, ion channels, o cell anchoring proteins, and cell type-specific surface antigens.
Many membrane proteins (MPs) contain amino acid sequence motifs that target these proteins to specific subcellular sites. Examples of these motifs include PDZ domains, KDEL, RGD, NGR, and GSL sequence motifs, von Willebrand factor A (vWFA) domains, and EGF-like domains. RGD, NGR, and GSL motif-containing peptides have been used as drug delivery agents in targeted cancer 5 treatment of tumor vasculature (Aap, W. et al. (1998) Science 279:377-380). Furthermore, MPs may also contain amino acid sequence motifs, such as the carbohydrate recognition domain (CRD), that mediate interactions with extracellular or intracellular molecules. G-Protein Coupled Receptors
G-protein coupled receptors (GPCR) are a superfamily of integral membrane proteins which o transduce extracellular signals. GPCRs include receptors for biogenic amines, lipid mediators of inflammation, peptide hormones, and sensory signal mediators. The structure of these highly-conserved receptors consists of seven hydrophobic transmembrane regions, an extracellular N-terminus, and a cytoplasmic C-terminus. Three extracellular loops alternate with three intracellular loops to link the seven transmembrane regions. Cysteine disulfide bridges connect the second and 5 third extracellular loops. The most conserved regions of GPCRs are the transmembrane regions and the first two cytoplasmic loops. A conserved, acidic- g-aromatic residue triplet present in the second cytoplasmic loop may interact with G proteins. A GPCR consensus pattern is characteristic of most proteins belonging to this superfamily (ExPASy PROSITE document PS00237; and Watson, S. and S. Akinstall (1994) The G-protein Linked Receptor Facts Book. Academic Press, San Diego CA, pp. 2-6). Mutations and changes in transcriptional activation of GPCR-encoding genes have been associated with neurological disorders such as schizophrenia, Parkinson's disease, Azheimer' s disease, drug addiction, and feeding disorders. Scavenger Receptors
Macrophage scavenger receptors with broad ligand specificity may participate in the binding of low density Upoproteins (LDL) and foreign antigens. Scavenger receptors types I and II are trimeric membrane proteins with each subunit containing a small N-terminal intracellular domain, a transmembrane domain, a large extracellular domain, and a C-terminal cysteine-rich domain. The extracellular domain contains a short spacer region, an α-helical coiled-coil region, and a triple helical collagen-like region. These receptors have been shown to bind a spectrum of ligands, including chemically modified Upoproteins and albumin, polyribonucleotides, polysaccharides, phospholipids, and asbestos (Matsumoto, A. et al. (1990) Proc. Natl. Acad. Sci. USA 87:9133-9137; and Elomaa, O. et al. (1995) Cell 80:603-609). The scavenger receptors are thought to play a key role in atherogenesis by mediating uptake of modified LDL in arterial walls, and in host defense by binding bacterial endotoxins, bacteria, and protozoa. Tetraspan Family Proteins
The transmembrane 4 superfamily (TM4SF) or tetraspan family is a multigene family encoding type III integral membrane proteins (Wright, M.D. and M.G. Tomlinson (1994) Immunol. Today 15 :588-594). The TM4SF is comprised of membrane proteins which traverse the cell membrane four times. Members of the TM4SF include platelet and endothelial cell membrane proteins, melanoma-associated antigens, leukocyte surface glycoproteins, colonal carcinoma antigens, tumor-associated antigens, and surface proteins of the schistosome parasites (Jankowski, S.A. (1994) Oncogene 9:1205-1211). Members of the TM4SF share about 25-30% amino acid sequence identity with one another.
A number of TM4SF members have been implicated in signal transduction, control of ceU adhesion, regulation of cell growth and proliferation, including development and oncogenesis, and cell motility, including tumor cell metastasis. Expression of TM4SF proteins is associated with a variety of tumors and the level of expression may be altered when cells are growing or activated. Tumor Antigens
Tumor antigens are cell surface molecules that are differentially expressed in tumor cells relative to normal cells. Tumor antigens distinguish tumor cells immunologically from normal cells and provide diagnostic and therapeutic targets for human cancers (Takagi, S. et al. (1995) Int. J. Cancer 61:706-715; Liu, E. et al. (1992) Oncogene 7:1027-1032). Leukocyte Antigens
Other types of cell surface antigens include those identified on leukocytic cells of the immune system. These antigens have been identified using systematic, monoclonal antibody (mAb)-based "shot gun" techniques. These techniques have resulted in the production of hundreds of mAbs directed against unknown cell surface leukocytic antigens. These antigens have been grouped into "clusters of differentiation" based on common immunocytochemical localization patterns in various differentiated and undifferentiated leukocytic cell types. Antigens in a given cluster are presumed to identify a single cell surface protein and are assigned a "cluster of differentiation" or "CD" designation. Some of the genes encoding proteins identified by CD antigens have been cloned and verified by standard molecular biology techniques. CD antigens have been characterized as both transmembrane proteins and cell surface proteins anchored to the plasma membrane via covalent attachment to fatty acid-containing glycolipids such as glycosylphosphatidylinositol (GPI). (Reviewed in Barclay, A.N. et al. (1995) The Leucocyte Antigen Facts Book. Academic Press, San Diego CA, pp. 17-20.) Ion Channels
Ion channels are found in the plasma membranes of virtually every cell in the body. For example, chloride channels mediate a variety of cellular functions including regulation of membrane potentials and absoφtion and secretion of ions across epithelial membranes. Chloride channels also regulate the pH of organelles such as the Golgi apparatus and endosomes (see, e.g., Greger, R. (1988) Annu. Rev. Physiol. 50:111-122). Electrophysiological and pharmacological properties of chloride channels, including ion conductance, current-voltage relationships, and sensitivity to modulators, suggest that different chloride channels exist in muscles, neurons, fibroblasts, epithelial cells, and lymphocytes.
Many ion channels have sites for phosphorylation by one or more protein kinases including protein kinase A, protein kinase C, tyrosine kinase, and casein kinase II, all of which regulate ion channel activity in cells. Inappropriate phosphorylation of proteins in cells has been Unked to changes in cell cycle progression and cell differentiation. Changes in the cell cycle have been linked to induction of apoptosis or cancer. Changes in cell differentiation have been linked to diseases and disorders of the reproductive system, immune system, skeletal muscle, and other organ systems. Proton Pumps
Proton ATPases comprise a large class of membrane proteins that use the energy of ATP hydrolysis to generate an electrochemical proton gradient across a membrane. The resultant gradient may be used to transport other ions across the membrane (Na+, K+, or Cl') or to maintain organelle pH. Proton ATPases are further subdivided into the mitochondrial F- ATPases, the plasma membrane ATPases, and the vacuolar ATPases. The vacuolar ATPases establish and maintain an acidic pH within various organelles involved in the processes of endocytosis and exocytosis (Mellman, I. et al. (1986) Annu. Rev. Biochem. 55:663-700). 5 Proton-coupled, 12 membrane-spanning domain transporters such as PEPT 1 and PEPT 2 are responsible for gastrointestinal absoφtion and for renal reabsoφtion of peptides using an electrochemical LT gradient as the driving force. Another type of peptide transporter, the TAP transporter, is a heterodimer consisting of TAP 1 and TAP 2 and is associated with antigen processing. Peptide antigens are transported across the membrane of the endoplasmic reticulum by l o TAP so they can be expressed on the cell surface in association with MHC molecules. Each TAP protein consists of multiple hydrophobic membrane spanning segments and a highly conserved ATP-binding cassette (Boll, M. et al. (1996) Proc. Natl. Acad. Sci. USA 93:284-289). Pathogenic microorganisms, such as heφes simplex virus, may encode inhibitors of TAP-mediated peptide transport in order to evade immune surveillance (Marusina, K. and J.J Manaco (1996) Curr. Opin.
15 Hematol. 3:19-26). ABC Transporters
The ATP-binding cassette (ABC) transporters, also called the "traffic ATPases", comprise a superfamily of membrane proteins that mediate transport and channel functions in prokaryotes and eukaryotes (Higgins, CF. (1992) Annu. Rev. Cell Biol. 8:67-113). ABC proteins share a similar
20 overall structure and significant sequence homology. All ABC proteins contain a conserved domain of approximately two hundred amino acid residues which includes one or more nucleotide binding domains. Mutations in ABC transporter genes are associated with various disorders, such as hyperbilirubinemia II/Dubin- Johnson syndrome, recessive Stargardt's disease, X-linked adrenoleukodystrophy, multidrug resistance, celiac disease, and cystic fibrosis.
25 Peripheral and Anchored Membrane Proteins
Some membrane proteins are not membrane-spanning but are attached to the plasma membrane via membrane anchors or interactions with integral membrane proteins. Membrane anchors are covalently joined to a protein post-ttanslationally and include such moieties as prenyl, myristyl, and glycosylphosphatidyl inositol groups. Membrane localization of peripheral and
3 o anchored proteins is important for their function in processes such as receptor-mediated signal transduction. For example, prenylation of Ras is required for its localization to the plasma membrane and for its normal and oncogenic functions in signal transduction. Vesicle Coat Proteins
Intercellular communication is essential for the development and survival of multicellular
35 organisms. Cells communicate with one another through the secretion and uptake of protein signaling molecules. The uptake of proteins into the cell is achieved by the endocytic pathway, in which the interaction of extracellular signaling molecules with plasma membrane receptors results in the formation of plasma membrane-derived vesicles that enclose and ttansport the molecules into the cytosol. These transport vesicles fuse with and mature into endosomal and lysosomal (digestive) 5 compartments. The secretion of proteins from the cell is achieved by exocytosis, in which molecules inside of the cell proceed through the secretory pathway. In this pathway, molecules transit from the ER to the Golgi apparatus and finally to the plasma membrane, where they are secreted from the cell. Several steps in the transit of material along the secretory and endocytic pathways require the formation of transport vesicles. Specifically, vesicles form at the transitional endoplasmic reticulum o (tER), the rim of Golgi cisternae, the face of the Trans-Golgi Network (TGN), the plasma membrane
(PM), and tubular extensions of the endosomes. Vesicle formation occurs when a region of membrane buds off from the donor organelle. The membrane-bound vesicle contains proteins to be transported and is surrounded by a proteinaceous coat, the components of which are recruited from the cytosol. Two different classes of coat protein have been identified. Clathrin coats form on 5 vesicles derived from the TGN and PM, whereas coatomer (COP) coats form on vesicles derived from the ER and Golgi. COP coats can be further classified as COPI, involved in retrograde traffic through the Golgi and from the Golgi to the ER, and COPII, involved in anterograde traffic from the ER to the Golgi (Mellman, supra).
In clathrin-based vesicle formation, adapter proteins bring vesicle cargo and coat proteins o together at the surface of the budding membrane. Adapter protein- 1 and -2 select cargo from the
TGN and plasma membrane, respectively, based on molecular information encoded on the cytoplasmic tail of integral membrane cargo proteins. Adapter proteins also recruit clathrin to the bud site. Clathrin is a protein complex consisting of three large and three small polypeptide chains arranged in a three-legged structure called a triskelion. Multiple triskelions and other coat proteins 5 appear to self -assemble on the membrane to form a coated pit. This assembly process may serve to deform the membrane into a budding vesicle. GTP-bound ADP-ribosylation factor (Arf) is also incoφorated into the coated assembly. Another small G-protein, dynamin, forms a ring complex around the neck of the forming vesicle and may provide the mechanochemical force to seal the bud, thereby releasing the vesicle. The coated vesicle complex is then transported through the cytosol. o During the transport process, Arf-bound GTP is hydrolyzed to GDP, and the coat dissociates from the transport vesicle (West, M.A. et al. (1997) J. Cell Biol. 138:1239-1254).
Vesicles which bud from the ER and the Golgi are covered with a protein coat similar to the clathrin coat of endocytic and TGN vesicles. The coat protein (COP) is assembled from cytosoUc precursor molecules at specific budding regions on the organelle. The COP coat consists of two 5 major components, a G-protein (Arf or Sar) and coat protomer (coatomer). Coatomer is an equimolar complex of seven proteins, termed alpha-, beta-, beta'-, gamma-, delta-, epsilon- and zeta-COP. The coatomer complex binds to dilysine motifs contained on the cytoplasmic tails of integral membrane proteins. These include the KKXX retrieval motif of membrane proteins of the ER and dibasic/diphenylamine motifs of members of the p24 family. The p24 family of type I membrane proteins represent the major membrane proteins of COPI vesicles (Harter, C and F.T. Wieland (1998) Proc. Natl. Acad. Sci. USA 95:11649-11654).
Organelle Associated Molecules
SEQ ID NO:54, SEQ ID NO:55, SEQ ID NO:56, SEQ ID NO:57, SEQ ID NO:58, SEQ ID NO:59, SEQ ID NO:60, SEQ ID NO:61, SEQ ID NO:62, and SEQ ID NO:63 encode, for example, organelle associated molecules.
Eukaryotic cells are organized into various cellular organelles which has the effect of separating specific molecules and their functions from one another and from the cytosol. Within the cell, various membrane structures surround and define these organelles while allowing them to interact with one another and the cell environment through both active and passive transport processes. Important cell organelles include the nucleus, the Golgi apparatus, the endoplasmic reticulum, mitochondria, peroxisomes, lysosomes, endosomes, and secretory vesicles. Nucleus
The cell nucleus contains all of the genetic information of the cell in the form of DNA, and the components and machinery necessary for replication of DNA and for transcription of DNA into RNA. (See Aberts, B. et al. (1994) Molecular Biology of the Cell, Garland Publishing Inc., New York NY, pp. 335-399.) DNA is organized into compact structures in the nucleus by interactions with various DNA-binding proteins such as histones and non-histone chromosomal proteins. DNA-specific nucleases, DNAses, partially degrade these compacted structures prior to DNA replication or transcription. DNA replication takes place with the aid of DNA helicases which unwind the double-stranded DNA helix, and DNA polymerases that duplicate the separated DNA strands.
Transcriptional regulatory proteins are essential for the control of gene expression. Some of these proteins function as transcription factors that initiate, activate, repress, or terminate gene transcription. Transcription factors generally bind to the promoter, enhancer, and upstream regulatory regions of a gene in a sequence-specific manner, although some factors bind regulatory elements within or downstream of a gene's coding region. Transcription factors may bind to a specific region of DNA singly or as a complex with other accessory factors. (Reviewed in Lewin, B. (1990) Genes IV, Oxford University Press, New York NY, and Cell Press, Cambridge MA, pp. 554-570.) Many transcription factors incoφorate DNA-binding structural motifs which comprise either α helices or β sheets that bind to the major groove of DNA. Four well-characterized structural motifs are helix-turn-heUx, zinc finger, leucine zipper, and helix-loop-helix. Proteins containing these motifs may act alone as monomers, or they may form homo- or heterodimers that interact with DNA. Many neoplastic disorders in humans can be attributed to inappropriate gene expression. Malignant cell growth may result from either excessive expression of tumor promoting genes or insufficient expression of tumor suppressor genes (Cleary, MX. (1992) Cancer Surv. 15:89-104). Chromosomal translocations may also produce chimeric loci which fuse the coding sequence of one gene with the regulatory regions of a second unrelated gene. Such an arrangement Ukely results in inappropriate gene transcription, potentially contributing to malignancy. In addition, the immune system responds to infection or trauma by activating a cascade of events that coordinate the progressive selection, amplification, and mobilization of cellular defense mechanisms. A complex and balanced program of gene activation and repression is involved in this process. However, hyperactivity of the immune system as a result of improper or insufficient regulation of gene expression may result in considerable tissue or organ damage. This damage is well documented in immunological responses associated with arthritis, allergens, heart attack, stroke, and infections (Isselbacher, KJ. et al. (1996) Harrison's Principles of Internal Medicine. 13/e, McGraw Hill, Inc. and Teton Data Systems Software).
Transcription of DNA into RNA also takes place in the nucleus catalyzed by RNA polymerases. Three types of RNA polymerase exist. RNA polymerase I makes large ribosomal RNAs, while RNA polymerase III makes a variety of small, stable RNAs including 5S ribosomal RNA and the transfer RNAs (tRNA). RNA polymerase II transcribes genes that will be translated into proteins. The primary transcript of RNA polymerase II is called heterogenous nuclear RNA (nnRNA), and must be further processed by splicing to remove non-coding sequences called introns. RNA splicing is mediated by small nuclear ribonucleoprotein complexes, or snRNPs, producing mature messenger RNA (mRNA) which is then transported out of the nucleus for translation into proteins. Nucleolus
The nucleolus is a highly organized subcompartment in the nucleus that contains high concentrations of RNA and proteins and functions mainly in ribosomal RNA synthesis and assembly (Alberts, et al. supra, pp. 379-382). Ribosomal RNA (rRNA) is a structural RNA that is complexed with proteins to form ribonucleoprotein structures called ribosomes. Ribosomes provide the platform on which protein synthesis takes place.
Ribosomes are assembled in the nucleolus initially from a large, 45S rRNA combined with a variety of proteins imported from the cytoplasm, as well as smaller, 5S rRNAs. Later processing of the immature ribosome results in formation of smaller ribosomal subunits which are transported from the nucleolus to the cytoplasm where they are assembled into functional ribosomes. Endoplasmic Reticulum
In eukaryotes, proteins are synthesized within the endoplasmic reticulum (ER), deUvered from the ER to the Golgi apparatus for post-translational processing and sorting, and transported from the 5 Golgi to specific intracellular and extracellular destinations. Synthesis of integral membrane proteins, secreted proteins, and proteins destined for the lumen of a particular organelle occurs on the rough endoplasmic reticulum (ER). The rough ER is so named because of the rough appearance in electron micrographs imparted by the attached ribosomes on which protein synthesis proceeds. Synthesis of proteins destined for the ER actually begins in the cytosol with the synthesis of a specific signal 0 peptide which directs the growing polypeptide and its attached ribosome to the ER membrane where the signal peptide is removed and protein synthesis is completed. Soluble proteins destined for the ER lumen, for secretion, or for transport to the lumen of other organelles pass completely into the ER lumen. Transmembrane proteins destined for the ER or for other cell membranes are translocated across the ER membrane but remain anchored in the lipid bilayer of the membrane by one or more 5 membrane-spanning α-helical regions.
Translocated polypeptide chains destined for other organelles or for secretion also fold and assemble in the ER lumen with the aid of certain "resident" ER proteins. Protein folding in the ER is aided by two principal types of protein isomerases, protein disulfide isomerase (PDI), and peptidyl- prolyl isomerase (PPI). PDI catalyzes the oxidation of free sulfhydryl groups in cysteine residues to o form intramolecular disulfide bonds in proteins. PPI, an enzyme that catalyzes the isomerization of certain proline imide bonds in oligopeptides and proteins, is considered to govern one of the rate limiting steps in the folding of many proteins to their final functional conformation. The cyclophilins represent a major class of PPI that was originally identified as the major receptor for the immunosuppressive drug cyclosporin A (Handschumacher, R.E. et al. (1984) Science 226:544-547). 5 Molecular "chaperones" such as BiP (binding protein) in the ER recognize incorrectly folded proteins as well as proteins not yet folded into their final form and bind to them, both to prevent improper aggregation between them, and to promote proper folding.
The "N-linked" glycosylation of most soluble secreted and membrane-bound proteins by oligosacchrides linked to asparagine residues in proteins is also performed in the ER. This reaction is o catalyzed by a membrane-bound enzyme, oligosaccharyl transferase.
Golgi Apparatus
The Golgi apparatus is a complex structure that lies adjacent to the ER in eukaryotic cells and serves primarily as a sorting and dispatching station for products of the ER (Aberts, et al. supra, pp. 600-610). Additional posttranslational processing, principally additional glycosylation, also occurs in the Golgi. Indeed, the Golgi is a major site of carbohydrate synthesis, including most of the glycosaminoglycans of the extracellular matrix. N-linked oUgosaccharides, added to proteins in the ER, are also further modified in the Golgi by the addition of more sugar residues to form complex N- linked oligosaccharides. "O-linked" glycosylation of proteins also occurs in the Golgi by the 5 addition of N-acetylgalactosamine to the hydroxyl group of a serine or threonine residue followed by the sequential addition of other sugar residues to the first. This process is catalyzed by a series of glycosyltransferases each specific for a particular donor sugar nucleotide and acceptor molecule (Lodish, H. et al. (1995) Molecular Cell Biology, W.H. Freeman and Co., New York NY, pp.700- 708). In many cases, both N- and O-linked oUgosaccharides appear to be required for the secretion of o proteins or the movement of plasma membrane glycoproteins to the cell surface.
The terminal compartment of the Golgi is the Trans-Golgi Network (TGN), where both membrane and lumenal proteins are sorted for their final destination. Transport (or secretory) vesicles destined for intracellular compartments, such as lysosomes, bud off of the TGN. Other transport vesicles bud off containing proteins destined for the plasma membrane, such as receptors, adhesion 5 molecules, and ion channels, and secretory proteins, such as hormones, neurotransmitters, and digestive enzymes. Vacuoles
The vacuole system is a collection of membrane bound compartments in eukaryotic cells that functions in the processes of endocytosis and exocytosis. They include phagosomes, lysosomes, o endosomes, and secretory vesicles. Endocytosis is the process in cells of internaUzing nutrients, solutes or small particles (pinocytosis) or large particles such as internaUzed receptors, viruses, bacteria, or bacterial toxins (phagocytosis). Exocytosis is the process of transporting molecules to the cell surface. It faciUtates placement or localization of membrane-bound receptors or other membrane proteins and secretion of hormones, neurotransmitters, digestive enzymes, wastes, etc. 5 A common property of all of these vacuoles is an acidic pH environment ranging from approximately pH 4.5-5.0. This acidity is maintained by the presence of a proton ATPase that uses the energy of ATP hydrolysis to generate an electrochemical proton gradient across a membrane (Mellman, I. et al. (1986) Annu. Rev. Biochem. 55:663-700). Eukaryotic vacuolar proton ATPase (vp-ATPase) is a multimeric enzyme composed of 3-10 different subunits. One of these subunits is a highly o hydrophobic polypeptide of approximately 16 kDa that is similar to the proteoUpid component of vp-
ATPases from eubacteria, fungi, and plant vacuoles (Mandel, M. et al. (1988) Proc. Natl. Acad. Sci. USA 85:5521-5524). The 16 kDa proteoUpid component is the major subunit of the membrane portion of vp-ATPase and functions in the transport of protons across the membrane. Lysosomes Lysosomes are membranous vesicles containing various hydrolytic enzymes used for the controlled intracellular digestion of macromolecules. Lysosomes contain some 40 types of enzymes including proteases, nucleases, glycosidases, Upases, phosphoUpases, phosphatases, and sulfatases, all of which are acid hydrolases that function at a pH of about 5. Lysosomes are surrounded by a unique 5 membrane containing ttansport proteins that allow the final products of macromolecule degradation, such as sugars, amino acids, and nucleotides, to be transported to the cytosol where they may be either excreted or reutilized by the cell. A vp-ATPase, such as that described above, maintains the acidic environment necessary for hydrolytic activity (Aberts, supra, pp. 610-611). Endosomes o Endosomes are another type of acidic vacuole that is used to transport substances from the cell surface to the interior of the cell in the process of endocytosis. Like lysosomes, endosomes have an acidic environment provided by a vp-ATPase (Alberts et al. supra, pp. 61 -618). Two types of endosomes are apparent based on tracer uptake studies that distinguish their time of formation in the cell and their cellular location. Early endosomes are found near the plasma membrane and appear to 5 function primarily in the recycling of intemaUzed receptors back to the cell surface. Late endosomes appear later in the endocytic process close to the Golgi apparatus and the nucleus, and appear to be associated with delivery of endocytosed material to lysosomes or to the TGN where they may be recycled. Specific proteins are associated with particular transport vesicles and their target compartments that may provide selectivity in targeting vesicles to their proper compartments. A o cytosoUc prenylated GTP-binding protein, Rab, is one such protein. Rabs 4, 5, and 11 are associated with the early endosome, whereas Rabs 7 and 9 associate with the late endosome. Mitochondria
Mitochondria are oval-shaped organelles comprising an outer membrane, a tightly folded inner membrane, an intermembrane space between the outer and inner membranes, and a matrix 5 inside the inner membrane. The outer membrane contains many porin molecules that allow ions and charged molecules to enter the intermembrane space, while the inner membrane contains a variety of transport proteins that transfer only selected molecules. Mitochondria are the primary sites of energy production in cells.
Energy is produced by the oxidation of glucose and fatty acids. Glucose is initially converted 0 to pyruvate in the cytoplasm. Fatty acids and pyruvate are transported to the mitochondria for complete oxidation to C02 coupled by enzymes to the transport of electrons from NADH and FADH2 to oxygen and to the synthesis of ATP (oxidative phosphorylation) from ADP and P;.
Pyruvate is transported into the mitochondria and converted to acetyl-CoA for oxidation via the citric acid cycle, involving pyruvate dehydrogenase components, dihydrolipoyl transacetylase, and dihydrolipoyl dehydrogenase. Enzymes involved in the citric acid cycle include: citrate synthetase, aconitases, isocitrate dehydrogenase, alpha-ketoglutarate dehydrogenase complex including transsuccinylases, succinyl CoA synthetase, succinate dehydrogenase, fumarases, and malate dehydrogenase. Acetyl CoA is oxidized to C02 with concomitant formation of NADH, FADH2, and GTP. In oxidative phosphorylation, the transfer of electrons from NADH and FADH2 to oxygen by dehydrogenases is coupled to the synthesis of ATP from ADP and P; by the F^Fλ ATPase complex in the mitochondrial inner membrane. Enzyme complexes responsible for electron transport and ATP synthesis include the FJ^ ATPase complex, ubiquinone(CoQ)-cytochrome c reductase, ubiquinone reductase, cytochrome b, cytochrome c1( FeS protein, and cytochrome c oxidase. Peroxisomes
Peroxisomes, like mitochondria, are a major site of oxygen utilization. They contain one or more enzymes, such as catalase and urate oxidase, that use molecular oxygen to remove hydrogen atoms from specific organic substrates in an oxidative reaction that produces hydrogen peroxide (Aberts, supra, pp. 574-577). Catalase oxidizes a variety of substrates including phenols, formic acid, formaldehyde, and alcohol and is important in peroxisomes of liver and kidney cells for detoxifying various toxic molecules that enter the bloodstream. Another major function of oxidative reactions in peroxisomes is the breakdown of fatty acids in a process called β oxidation, β oxidation results in shortening of the alkyl chain of fatty acids by blocks of two carbon atoms that are converted to acetyl CoA and exported to the cytosol for reuse in biosynthetic reactions. Aso like mitochondria, peroxisomes import their proteins from the cytosol using a specific signal sequence located near the C-terminus of the protein. The importance of this import process is evident in the inherited human disease Zellweger syndrome, in which a defect in importing proteins into perixosomes leads to a perixosomal deficiency resulting in severe abnormalities in the brain, liver, and kidneys, and death soon after birth. One form of this disease has been shown to be due to a mutation in the gene encoding a perixosomal integral membrane protein called peroxisome assembly factor- 1.
The discovery of new human molecules satisfies a need in the art by providing new compositions which are useful in the diagnosis, study, prevention, and treatment of diseases associated with, as well as effects of exogenous compounds on, the expression of human molecules.
SUMMARY OF THE INVENTION
The present invention relates to nucleic acid sequences comprising human diagnostic and therapeutic polynucleotides (dithp) as presented in the Sequence Listing. Some of the dithp uniquely identify genes encoding human structural, functional, and regulatory molecules. The invention provides an isolated polynucleotide comprising a polynucleotide sequence selected from the group consisting of a) a polynucleotide sequence selected from the group consisting of SEQ ID NO: 1-71 ; b) a naturally occurring polynucleotide sequence having at least 90% sequence identity to a polynucleotide sequence selected from the group consisting of SEQ ID NO: 1-71 ; c) a polynucleotide sequence complementary to a); d) a polynucleotide sequence complementary to b); and e) an RNA equivalent of a) through d). In one alternative, the polynucleotide comprises a polynucleotide sequence selected from the group consisting of SEQ ID NO: 1-71. In another alternative, the polynucleotide comprises at least 60 contiguous nucleotides of a polynucleotide sequence selected from the group consisting of a) a polynucleotide sequence selected from the group consisting of SEQ ID NO: 1 -71 ; b) a naturally occurring polynucleotide sequence having at least 90% sequence identity to a polynucleotide sequence selected from the group consisting of SEQ ID NO: 1-71 ; c) a polynucleotide sequence complementary to a); d) a polynucleotide sequence complementary to b); and e) an RNA equivalent of a) through d). The invention further provides a composition for the detection of expression of human diagnostic and therapeutic polynucleotides, comprising at least one isolated polynucleotide comprising a polynucleotide sequence selected from the group consisting of a) a polynucleotide sequence selected from the group consisting of SEQ ID NO: 1-71 ; b) a naturally occurring polynucleotide sequence having at least 90% sequence identity to a polynucleotide sequence selected from the group consisting of SEQ ID NO: 1-71 ; c) a polynucleotide sequence complementary to a); d) a polynucleotide sequence complementary to b); and e) an RNA equivalent of a) through d); and a detectable label.
The invention also provides a method for detecting a target polynucleotide in a sample, said target polynucleotide comprising a polynucleotide sequence selected from the group consisting of a) a polynucleotide sequence selected from the group consisting of SEQ ID NO: 1-71; b) a naturally occurring polynucleotide sequence having at least 90% sequence identity to a polynucleotide sequence selected from the group consisting of SEQ ID NO: 1-71 ; c) a polynucleotide sequence complementary to a); d) a polynucleotide sequence complementary to b); and e) an RNA equivalent of a) through d). The method comprises a) hybridizing the sample with a probe comprising at least 20 contiguous nucleotides comprising a sequence complementary to said target polynucleotide in the sample, and which probe specifically hybridizes to said target polynucleotide, under conditions whereby a hybridization complex is formed between said probe and said target polynucleotide, and b) detecting the presence or absence of said hybridization complex, and, optionally, if present, the amount thereof. In one alternative, the probe comprises at least 30 contiguous nucleotides. In another alternative, the probe comprises at least 60 contiguous nucleotides.
The invention further provides a recombinant polynucleotide comprising a promoter sequence operably Unked to an isolated polynucleotide comprising a polynucleotide sequence selected from the group consisting of a) a polynucleotide sequence selected from the group consisting of SEQ ID NO:l- 71 ; b) a naturally occurring polynucleotide sequence having at least 90% sequence identity to a polynucleotide sequence selected from the group consisting of SEQ ID NO: 1-71 ; c) a polynucleotide 5 sequence complementary to a); d) a polynucleotide sequence complementary to b); and e) an RNA equivalent of a) through d). In one alternative, the invention provides a cell transformed with the recombinant polynucleotide. In another alternative, the invention provides a transgenic organism comprising the recombinant polynucleotide. In a further alternative, the invention provides a method for producing a human diagnostic and therapeutic polypeptide, the method comprising a) culturing a o cell under conditions suitable for expression of the human diagnostic and therapeutic polypeptide, wherein said cell is transformed with the recombinant polynucleotide, and b) recovering the human diagnostic and therapeutic polypeptide so expressed.
The invention also provides a purified human diagnostic and therapeutic polypeptide (DITHP) encoded by at least one polynucleotide comprising a polynucleotide sequence selected from the group 5 consisting of SEQ ID NO:l-71. Additionally, the invention provides an isolated antibody which specifically binds to the human diagnostic and therapeutic polypeptide. The invention further provides a method of identifying a test compound which specifically binds to the human diagnostic and therapeutic polypeptide, the method comprising the steps of a) providing a test compound; b) combining the human diagnostic and therapeutic polypeptide with the test compound for a sufficient time and o under suitable conditions for binding; and c) detecting binding of the human diagnostic and therapeutic polypeptide to the test compound, thereby identifying the test compound which specifically binds the human diagnostic and therapeutic polypeptide.
The invention further provides a microarray wherein at least one element of the microarray is an isolated polynucleotide comprising at least 60 contiguous nucleotides of a polynucleotide comprising 5 a polynucleotide sequence selected from the group consisting of a) a polynucleotide sequence selected from the group consisting of SEQ ID NO: 1 -71 ; b) a naturally occurring polynucleotide sequence having at least 90% sequence identity to a polynucleotide sequence selected from the group consisting of SEQ ID NO:l-71 ; c) a polynucleotide sequence complementary to a); d) a polynucleotide sequence complementary to b); and e) an RNA equivalent of a) through d). The invention also provides a method o for generating a transcript image of a sample which contains polynucleotides. The method comprises a) labeling the polynucleotides of the sample, b) contacting the elements of the microarray with the labeled polynucleotides of the sample under conditions suitable for the formation of a hybridization complex, and c) quantifying the expression of the polynucleotides in the sample.
Additionally, the invention provides a method for screening a compound for effectiveness in altering expression of a target polynucleotide, wherein said target polynucleotide comprises a polynucleotide sequence selected from the group consisting of a) a polynucleotide sequence selected from the group consisting of SEQ ID NO: 1-71 ; b) a naturally occurring polynucleotide sequence having at least 90% sequence identity to a polynucleotide sequence selected from the group consisting of SEQ 5 ID NO:l-71 ; c) a polynucleotide sequence complementary to a); d) a polynucleotide sequence complementary to b); and e) an RNA equivalent of a) through d). The method comprises a) exposing a sample comprising the target polynucleotide to a compound, and b) detecting altered expression of the target polynucleotide.
The invention further provides a method for assessing toxicity of a test compound, said method 0 comprising a) treating a biological sample containing nucleic acids with the test compound; b) hybridizing the nucleic acids of the treated biological sample with a probe comprising at least 20 contiguous nucleotides of a polynucleotide comprising a polynucleotide sequence selected from the group consisting of i) a polynucleotide sequence selected from the group consisting of SEQ ID NO:l- 71 ; U) a naturally occurring polynucleotide sequence having at least 90% sequence identity to a 5 polynucleotide sequence selected from the group consisting of SEQ ID NO : 1 -71 ; Ui) a polynucleotide sequence complementary to i), iv) a polynucleotide sequence complementary to u), and v) an RNA equivalent of i)-iv). Hybridization occurs under conditions whereby a specific hybridization complex is formed between said probe and a target polynucleotide in the biological sample, said target polynucleotide comprising a polynucleotide sequence selected from the group consisting of i) a o polynucleotide sequence selected from the group consisting of SEQ ID NO: 1 -71 ; ii) a naturally occurring polynucleotide sequence having at least 90% sequence identity to a polynucleotide sequence selected from the group consisting of SEQ ID NO: 1-71; iii) a polynucleotide sequence complementary to i), iv) a polynucleotide sequence complementary to ii), and v) an RNA equivalent of i)-iv), and alternatively, the target polynucleotide comprises a fragment of a polynucleotide sequence 5 selected from the group consisting of i-v above; c) quantifying the amount of hybridization complex; and d) comparing the amount of hybridization complex in the treated biological sample with the amount of hybridization complex in an untreated biological sample, wherein a difference in the amount of hybridization complex in the treated biological sample is indicative of toxicity of the test compound.
o DESCRIPTION OF THE TABLES
Table 1 shows the sequence identification numbers (SEQ ID NO:s) and template identification numbers (template IDs) corresponding to the polynucleotides of the present invention, along with their GenBank hits (GI Numbers), probabiUty scores, and functional annotations corresponding to the GenBank hits. Table 2 shows the sequence identification numbers (SEQ ID NO:s) and template identification numbers (template IDs) corresponding to the polynucleotides of the present invention, along with polynucleotide segments of each template sequence as defined by the indicated "start" and "stop" nucleotide positions. The reading frames of the polynucleotide segments and the Pfam hits, Pfam 5 descriptions, and E- values corresponding to the polypeptide domains encoded by the polynucleotide segments are indicated.
Table 3 shows the sequence identification numbers (SEQ ID NO:s) and template identification numbers (template IDs) corresponding to the polynucleotides of the present invention, along with polynucleotide segments of each template sequence as defined by the indicated "start" and "stop" o nucleotide positions. The reading frames of the polynucleotide segments are shown, and the polypeptides encoded by the polynucleotide segments constitute either signal peptide (SP) or transmembrane (TM) domains, as indicated.
Table 4 shows the sequence identification numbers (SEQ ID NO:s) and template identification numbers (template IDs) corresponding to the polynucleotides of the present invention, along with 5 component sequence identification numbers (component IDs) corresponding to each template. The component sequences, which were used to assemble the template sequences, are defined by the indicated "start" and "stop" nucleotide positions along each template.
Table 5 shows the tissue distribution profiles for the templates of the invention.
Table 6 summarizes the bioinformatics tools which are useful for analysis of the o polynucleotides of the present invention. The first column of Table 6 Usts analytical tools, programs, and algorithms, the second column provides brief descriptions thereof, the third column presents appropriate references, all of which are incoφorated by reference herein in their entirety, and the fourth column presents, where appUcable, the scores, probabiUty values, and other parameters used to evaluate the strength of a match between two sequences (the higher the score, the greater the homology between 5 two sequences).
DETAILED DESCRIPTION OF THE INVENTION
Before the nucleic acid sequences and methods are presented, it is to be understood that this invention is not Umited to the particular machines, methods, and materials described. Athough o particular embodiments are described, machines, methods, and materials similar or equivalent to these embodiments may be used to practice the invention. The preferred machines, methods, and materials set forth are not intended to Umit the scope of the invention which is Umited only by the appended claims.
The singular forms "a", "an", and "the" include plural reference unless the context clearly dictates otherwise. Al technical and scientific terms have the meanings commonly understood by one of ordinary skill in the art. Al publications are incoφorated by reference for the puφose of describing and disclosing the cell Unes, vectors, and methodologies which are presented and which might be used in connection with the invention. Nothing in the specification is to be construed as an admission that the 5 invention is not entitled to antedate such disclosure by virtue of prior invention.
Definitions
As used herein, the lower case "dithp" refers to a nucleic acid sequence, while the upper case "DITHP" refers to an amino acid sequence encoded by dithp. A "full-length" dithp refers to a nucleic o acid sequence containing the entire coding region of a gene endogenously expressed in human tissue.
"Adjuvants" are materials such as Freund's adjuvant, mineral gels (aluminum hydroxide), and surface active substances (lysolecitbin, pluronic polyols, polyanions, peptides, oil emulsions, keyhole limpet hemocyanin, and dinifrophenol) which may be administered to increase a host's immunological response. 5 " Alele" refers to an alternative form of a nucleic acid sequence. Aleles result from a
"mutation," a change or an alternative reading of the genetic code. A y given gene may have none, one, or many allelic forms. Mutations which give rise to alleles include deletions, additions, or substitutions of nucleotides. Each of these changes may occur alone, or in combination with the others, one or more times in a given nucleic acid sequence. The present invention encompasses alleUc dithp. 0 "Anino acid sequence" refers to a peptide, a polypeptide, or a protein of either natural or synthetic origin. The amino acid sequence is not Umited to the complete, endogenous amino acid sequence and may be a fragment, epitope, variant, or derivative of a protein expressed by a nucleic acid sequence.
"AnpUfication" refers to the production of additional copies of a sequence and is carried out 5 using polymerase chain reaction (PCR) technologies well known in the art.
"Aitibody" refers to intact molecules as well as to fragments thereof, such as Fab, F(ab')2, and Fv fragments, which are capable of binding the epitopic determinant. Aitibodies that bind DITHP polypeptides can be prepared using intact polypeptides or using fragments containing small peptides of interest as the immunizing antigen. The polypeptide or peptide used to immunize an animal (e.g., a o mouse, a rat, or a rabbit) can be derived from the translation of RNA, or synthesized chemically, and can be conjugated to a carrier protein if desired. Commonly used carriers that are chemically coupled to peptides include bovine serum albumin, thyroglobulin, and keyhole Umpet hemocyanin (KLH). The coupled peptide is then used to immunize the animal.
"Antisense sequence" refers to a sequence capable of specifically hybridizing to a target sequence. The antisense sequence may include DNA, RNA, or any nucleic acid mimic or analog such as peptide nucleic acid (PNA); oUgonucleotides having modified backbone Unkages such as phosphorothioates, methylphosphonates, or benzylphosphonates; oUgonucleotides having modified sugar groups such as 2 -methoxyethyl sugars or 2 -methoxyethoxy sugars; or oligonucleotides having modified bases such as 5-methyl cytosine, 2 -deoxyuracil, or 7-deaza-2'-deoxyguanosine.
"Antisense sequence" refers to a sequence capable of specifically hybridizing to a target sequence. The antisense sequence can be DNA, RNA, or any nucleic acid mimic or analog.
"Antisense technology" refers to any technology which reUes on the specific hybridization of an antisense sequence to a target sequence. A "bin" is a portion of computer memory space used by a computer program for storage of data, and bounded in such a manner that data stored in a bin may be retrieved by the program.
"Biologically active" refers to an amino acid sequence having a structural, regulatory, or biochemical function of a naturally occurring amino acid sequence.
"Clone joining" is a process for combining gene bins based upon the bins' containing sequence information from the same clone. The sequences may assemble into a primary gene transcript as well as one or more spUce variants.
"Complementary" describes the relationship between two single-stranded nucleic acid sequences that anneal by base-pairing (5'-A-G-T-3' pairs with its complement 3'-T-C-A-5').
A "component sequence" is a nucleic acid sequence selected by a computer program such as PHRED and used to assemble a consensus or template sequence from one or more component sequences.
A "consensus sequence" or "template sequence" is a nucleic acid sequence which has been assembled from overlapping sequences, using a computer program for fragment assembly such as the GEL VIEW fragment assembly system (Genetics Computer Group (GCG), Madison WI) or using a relational database management system (RDMS).
"Conservative amino acid substitutions" are those substitutions that, when made, least interfere with the properties of the original protein, i.e., the structure and especially the function of the protein is conserved and not significantly changed by such substitutions. The table below shows amino acids which may be substituted for an original amino acid in a protein and which are regarded as conservative substitutions.
Original Residue Conservative Substitution
Aa Gly, Ser
Ag His, Lys Asn Asp, Gin, His Asp An, Glu
Cys Aa, Ser
Gin An, Glu, His
Glu Ap, Gin, His
Gly Aa
His An, Ag, Gin, Glu lie Leu, Val
Leu lie, Val
Lys Ag, Gin, Glu
Met Leu, He
Phe His, Met, Leu, Tφ, Tyr
Ser Cys, Thr
Thr Ser, Val
Tφ Phe, Tyr
Tyr His, Phe, Tφ
Val lie, Leu, Thr
Conservative substitutions generally maintain (a) the structure of the polypeptide backbone in the area of the substitution, for example, as a beta sheet or alpha hetical conformation, (b) the charge or hydrophobicity of the molecule at the target site, or (c) the bulk of the side chain.
"Deletion" refers to a change in either a nucleic or amino acid sequence in which at least one nucleotide or amino acid residue, respectively, is absent.
"Derivative" refers to the chemical modification of a nucleic acid sequence, such as by replacement of hydrogen by an alkyl, acyl, amino, hydroxyl, or other group.
The terms "element" and "array element" refer to a polynucleotide, polypeptide, or other chemical compound having a unique and defined position on a microarray.
'Ε- value" refers to the statistical probabiUty that a match between two sequences occurred by chance. A "fragment" is a unique portion of dithp or DITHP which is identical in sequence to but shorter in length than the parent sequence. A fragment may comprise up to the entire length of the defined sequence, minus one nucleotide/amino acid residue. For example, a fragment may comprise from 10 to 1000 contiguous amino acid residues or nucleotides. A fragment used as a probe, primer, antigen, therapeutic molecule, or for other puφoses, may be at least 5, 10, 15, 16, 20, 25, 30, 40, 50, 60, 75, 100, 150, 250 or at least 500 contiguous amino acid residues or nucleotides in length.
Fragments may be preferentially selected from certain regions of a molecule. For example, a polypeptide fragment may comprise a certain length of contiguous amino acids selected from the first 250 or 500 amino acids (or first 25% or 50%) of a polypeptide as shown in a certain defined sequence. Clearly these lengths are exemplary, and any length that is supported by the specification, including the Sequence Listing and the figures, may be encompassed by the present embodiments.
A fragment of dithp comprises a region of unique polynucleotide sequence that specifically identifies dithp, for example, as distinct from any other sequence in the same genome. A fragment of 5 dithp is useful, for example, in hybridization and amplification technologies and in analogous methods that distinguish dithp from related polynucleotide sequences. The precise length of a fragment of dithp and the region of dithp to which the fragment corresponds are routinely determinable by one of ordinary skill in the art based on the intended puφose for the fragment.
A fragment of DITHP is encoded by a fragment of dithp. A fragment of DITHP comprises a 0 region of unique amino acid sequence that specifically identifies DITHP. For example, a fragment of DITHP is useful as an immunogenic peptide for the development of antibodies that specifically recognize DITHP. The precise length of a fragment of DITHP and the region of DITHP to which the fragment corresponds are routinely determinable by one of ordinary skill in the art based on the intended puφose for the fragment. 5 A "full length" nucleotide sequence is one containing at least a start site for translation to a protein sequence, followed by an open reading frame and a stop site, and encoding a "full length" polypeptide.
"Hit" refers to a sequence whose annotation will be used to describe a given template. Criteria for selecting the top hit are as follows: if the template has one or more exact nucleic acid matches, the o top hit is the exact match with highest percent identity. If the template has no exact matches but has significant protein hits, the top hit is the protein hit with the lowest E-value. If the template has no significant protein hits, but does have significant non-exact nucleotide hits, the top hit is the nucleotide hit with the lowest E-value.
"Homology" refers to sequence similarity either between a reference nucleic acid sequence and 5 at least a fragment of a dithp or between a reference amino acid sequence and a fragment of a DITHP.
"Hybridization" refers to the process by which a strand of nucleotides anneals with a complementary strand through base pairing. Specific hybridization is an indication that two nucleic acid sequences share a high degree of identity. Specific hybridization complexes form under defined annealing conditions, and remain hybridized after the "washing" step. The defined hybridization o conditions include the anneahng conditions and the washing step(s), the latter of which is particularly important in determining the stringency of the hybridization process, with more stringent conditions allowing less non-specific binding, i.e., binding between pairs of nucleic acid probes that are not perfectly matched. Permissive conditions for anneahng of nucleic acid sequences are routinely determinable and may be consistent among hybridization experiments, whereas wash conditions may be varied among experiments to achieve the desired stringency.
Generally, stringency of hybridization is expressed with reference to the temperature under which the wash step is carried out. Generally, such wash temperatures are selected to be about 5°C to 5 20°C lower than the thermal melting point (TJ for the specific sequence at a defined ionic strength and pH. The Tm is the temperature (under defined ionic strength and pH) at which 50% of the target sequence hybridizes to a perfectly matched probe. An equation for calculating Tm and conditions for nucleic acid hybridization is well known and can be found in Sambrook et al., 1989, Molecular Cloning: A Laboratory Manual. 2nd ed., vol. 1-3, Cold Spring Harbor Press, Plainview NY; specifically o see volume 2, chapter 9.
High stringency conditions for hybridization between polynucleotides of the present invention include wash conditions of 68°C in the presence of about 0.2 x SSC and about 0.1 % SDS, for 1 hour. Aternatively, temperatures of about 65°C, 60°C, or 55°C may be used. SSC concentration may be varied from about 0.2 to 2 x SSC, with SDS being present at about 0.1 %. Typically, blocking reagents 5 are used to block non-specific hybridizatioa Such blocking reagents include, for instance, denatured salmon sperm DNA at about 100-200 μg/ml. Useful variations on these conditions will be readily apparent to those skilled in the art. Hybridization, particularly under high stringency conditions, may be suggestive of evolutionary similarity between the nucleotides. Such similarity is strongly indicative of a similar role for the nucleotides and their resultant proteins. o Other parameters, such as temperature, salt concentration, and detergent concentration may be varied to achieve the desired stringency. Denaturants, such as formamide at a concentration of about 35-50% v/v, may also be used under particular circumstances, such as RNA:DNA hybridizations. Appropriate hybridization conditions are routinely determinable by one of ordinary skill in the art.
"Immunogenic" describes the potential for a natural, recombinant, or synthetic peptide, epitope, 5 polypeptide, or protein to induce antibody production in appropriate animals, cells, or cell lines.
"Insertion" or "addition" refers to a change in either a nucleic or amino acid sequence in which at least one nucleotide or residue, respectively, is added to the sequence.
"LabeUng" refers to the covalent or noncovalent joining of a polynucleotide, polypeptide, or antibody with a reporter molecule capable of producing a detectable or measurable signal. o "Microarray" is any arrangement of nucleic acids, amino acids, antibodies, etc., on a substrate.
The substrate may be a sohd support such as beads, glass, paper, nitrocellulose, nylon, or an appropriate membrane.
"Linkers" are short stretches of nucleotide sequence which may be added to a vector or a dithp to create restriction endonuclease sites to faciUtate cloning. "PolyUnkers" are engineered to incoφorate multiple restriction enzyme sites and to provide for the use of enzymes which leave 5 ' or 3 ' overhangs (e.g., BamHI, EcoRI, and Hindlll) and those which provide blunt ends (e.g., EcoRV, SnaBI, and Stul).
"Naturally occurring" refers to an endogenous polynucleotide or polypeptide that may be isolated from viruses or prokaryotic or eukaryotic cells. 5 "Nucleic acid sequence" refers to the specific order of nucleotides joined by phosphodiester bonds in a linear, polymeric arrangement. Depending on the number of nucleotides, the nucleic acid sequence can be considered an oUgomer, oligonucleotide, or polynucleotide. The nucleic acid can be DNA, RNA, or any nucleic acid analog, such as PNA, may be of genomic or synthetic origin, may be either double-stranded or single-stranded, and can represent either the sense or antisense 0 (complementary) strand.
"OUgomer" refers to a nucleic acid sequence of at least about 6 nucleotides and as many as about 60 nucleotides, preferably about 15 to 40 nucleotides, and most preferably between about 20 and 30 nucleotides, that may be used in hybridization or ampUfication technologies. OUgomers may be used as, e.g., primers for PCR, and are usually chemically synthesized. 5 "Operably linked" refers to the situation in which a first nucleic acid sequence is placed in a functional relationship with the second nucleic acid sequence. For instance, a promoter is operably Unked to a coding sequence if the promoter affects the transcription or expression of the coding sequence. Generally, operably Unked DNA sequences may be in close proximity or contiguous and, where necessary to join two protein coding regions, in the same reading frame. o "Peptide nucleic acid" (PNA) refers to a DNA mimic in which nucleotide bases are attached to a pseudopeptide backbone to increase stability. PNA, also designated antigene agents, can prevent gene expression by targeting complementary messenger RNA.
The phrases "percent identity" and "% identity", as apphed to polynucleotide sequences, refer to the percentage of residue matches between at least two polynucleotide sequences aUgned using a 5 standardized algorithm. Such an algorithm may insert, in a standardized and reproducible way, gaps in the sequences being compared in order to optimize aUgnment between two sequences, and therefore achieve a more meaningful comparison of the two sequences.
Percent identity between polynucleotide sequences may be determined using the default parameters of the CLUSTAL V algorithm as incoφorated into the MEGALIGN version 3.12e sequence o aUgnment program. This program is part of the LASERGENE software package, a suite of molecular biological analysis programs (DNASTAR, Madison WI). CLUSTAL V is described in Higgins, D.G. and Shaφ, P.M. (1989) CABIOS 5:151-153 and in Higgins, D.G. et al. (1992) CABIOS 8:189-191. For pairwise aUgnments of polynucleotide sequences, the default parameters are set as follows: Ktuple=2, gap penalty=5, window=4, and "diagonals saved"=4. The "weighted" residue weight table is selected as the default. Percent identity is reported by CLUSTAL V as the "percent similarity" between aUgned polynucleotide sequence pairs.
Aternatively, a suite of commonly used and freely available sequence comparison algorithms is provided by the National Center for Biotechnology Information (NCBI) Basic Local AUgnment Search 5 Tool (BLAST) (Atschul, S.F. et al. (1990) J. Mol. Biol. 215:403-410), which is available from several sources, including the NCBI, Bethesda, MD, and on the Internet at http://www.ncbi.nlm.nih.gov/BLAST/. The BLAST software suite includes various sequence analysis programs including "blastn," that is used to determine aUgnment between a known polynucleotide sequence and other sequences on a variety of databases. Aso available is a tool called "BLAST 2 o Sequences" that is used for direct pairwise comparison of two nucleotide sequences. "BLAST 2
Sequences" can be accessed and used interactively at http://www.ncbi.nlm.nih.gov/gorf/bl2/. The "BLAST 2 Sequences" tool can be used for both blastn and blastp (discussed below). BLAST programs are commonly used with gap and other parameters set to default settings. For example, to compare two nucleotide sequences, one may use blastn with the "BLAST 2 Sequences" tool Version 5 2.0.9 (May-07-1999) set at default parameters. Such default parameters may be, for example:
Matrix: BLOSUM62
Reward for match: 1
Penalty for mismatch: -2
Open Gap: 5 and Extension Gap: 2 penalties o Gap x drop-off: 50
Expect: 10
Word Size: 11
Filter: on
Percent identity may be measured over the length of an entire defined sequence, for example, as 5 defined by a particular SEQ ID number, or may be measured over a shorter length, for example, over the length of a fragment taken from a larger, defined sequence, for instance, a fragment of at least 20, at least 30, at least 40, at least 50, at least 70, at least 100, or at least 200 contiguous nucleotides. Such lengths are exemplary only, and it is understood that any fragment length supported by the sequences shown herein, in figures or Sequence Listings, may be used to describe a length over which percentage o identity may be measured.
Nucleic acid sequences that do not show a high degree of identity may nevertheless encode similar amino acid sequences due to the degeneracy of the genetic code. It is understood that changes in nucleic acid sequence can be made using this degeneracy to produce multiple nucleic acid sequences that all encode substantially the same proteia The phrases "percent identity" and "% identity", as appUed to polypeptide sequences, refer to the percentage of residue matches between at least two polypeptide sequences aUgned using a standardized algorithm. Methods of polypeptide sequence aUgnment are well-known. Some aUgnment methods take into account conservative amino acid substitutions. Such conservative substitutions, explained in more detail above, generally preserve the hydrophobicity and acidity of the substituted residue, thus preserving the structure (and therefore function) of the folded polypeptide.
Percent identity between polypeptide sequences may be determined using the default parameters of the CLUSTAL V algorithm as incoφorated into the MEGALIGN version 3.12e sequence alignment program (described and referenced above). For pairwise alignments of polypeptide sequences using CLUSTAL V, the default parameters are set as follows: Ktuple=l, gap penalty=3, window=5, and "diagonals saved"=5. The PAM250 matrix is selected as the default residue weight table. As with polynucleotide aUgnments, the percent identity is reported by CLUSTAL V as the "percent similarity" between aUgned polypeptide sequence pairs.
Aternatively the NCBI BLAST software suite may be used. For example, for a pairwise comparison of two polypeptide sequences, one may use the "BLAST 2 Sequences" tool Version 2.0.9 (May-07-1999) with blastp set at default parameters. Such default parameters may be, for example:
Matrix: BLOSUM62
Open Gap: 11 and Extension Gap: 1 penalty
Gap x drop-off: 50 Expect: 10
Word Size: 3
Filter: on
Percent identity may be measured over the length of an entire defined polypeptide sequence, for example, as defined by a particular SEQ ID number, or may be measured over a shorter length, for example, over the length of a fragment taken from a larger, defined polypeptide sequence, for instance, a fragment of at least 15, at least 20, at least 30, at least 40, at least 50, at least 70 or at least 150 contiguous residues. Such lengths are exemplary only, and it is understood that any fragment length supported by the sequences shown herein, in figures or Sequence Listings, may be used to describe a length over which percentage identity may be measured. "Post-translational modification" of a DITHP may involve Upidation, glycosylation, phosphorylation, acetylation, racemization, proteolytic cleavage, and other modifications known in the art. These processes may occur synthetically or biochemically. Biochemical modifications will vary by cell type depending on the enzymatic miUeu and the DITHP. "Probe" refers to dithp or fragments thereof, which are used to detect identical, alletic or related nucleic acid sequences. Probes are isolated oUgonucleotides or polynucleotides attached to a detectable label or reporter molecule. Typical labels include radioactive isotopes, Ugands, chemiluminescent agents, and enzymes. "Primers" are short nucleic acids, usually DNA oUgonucleotides, which may be 5 annealed to a target polynucleotide by complementary base-pairing. The primer may then be extended along the target DNA strand by a DNA polymerase enzyme. Primer pairs can be used for ampUfication (and identification) of a nucleic acid sequence, e.g., by the polymerase chain reaction (PCR).
Probes and primers as used in the present invention typically comprise at least 15 contiguous nucleotides of a known sequence. In order to enhance specificity, longer probes and primers may also 0 be employed, such as probes and primers that comprise at least 20, 30, 40, 50, 60, 70, 80, 90, 100, or at least 150 consecutive nucleotides of the disclosed nucleic acid sequences. Probes and primers may be considerably longer than these examples, and it is understood that any length supported by the specification, including the figures and Sequence Listing, may be used.
Methods for preparing and using probes and primers are described in the references, for 5 example Sambrook et al., 1989, Molecular Cloning: A Laboratory Manual, 2nd ed., vol. 1-3, Cold Spring Harbor Press, Plainview NY; Ausubel et al.,1987, Current Protocols in Molecular Biology, Greene Publ. Asoc. & Wiley-Intersciences, New York NY; Innis et al., 1990, PCR Protocols, A Guide to Methods and AppUcations, Academic Press, San Diego CA. PCR primer pairs can be derived from a known sequence, for example, by using computer programs intended for that puφose such as Primer o (Version 0.5, 1991 , Whitehead Institute for Biomedical Research, Cambridge MA).
OUgonucleotides for use as primers are selected using software known in the art for such puφose. For example, OLIGO 4.06 software is useful for the selection of PCR primer pairs of up to 100 nucleotides each, and for the analysis of oUgonucleotides and larger polynucleotides of up to 5,000 nucleotides from an input polynucleotide sequence of up to 32 kilobases. Similar primer selection 5 programs have incoφorated additional features for expanded capabiUties. For example, the PrimOU primer selection program (available to the pubUc from the Genome Center at University of Texas South West Medical Center, Dallas TX) is capable of choosing specific primers from megabase sequences and is thus useful for designing primers on a genome-wide scope. The Primer3 primer selection program (available to the pubUc from the Whitehead Institute/MIT Center for Genome Research, o Cambridge MA) allows the user to input a "mispriming Ubrary," in which sequences to avoid as primer binding sites are user-specified. Primer3 is useful, in particular, for the selection of oUgonucleotides for microarrays. (The source code for the latter two primer selection programs may also be obtained from their respective sources and modified to meet the user's specific needs.) The PrimeGen program (available to the public from the UK Human Genome Mapping Project Resource Centre, Cambridge UK) designs primers based on multiple sequence alignments, thereby allowing selection of primers that hybridize to either the most conserved or least conserved regions of aUgned nucleic acid sequences.
Hence, this program is useful for identification of both unique and conserved oligonucleotides and polynucleotide fragments. The oUgonucleotides and polynucleotide fragments identified by any of the above selection methods are useful in hybridization technologies, for example, as PCR or sequencing primers, microarray elements, or specific probes to identify fully or partially complementary polynucleotides in a sample of nucleic acids. Methods of oUgonucleotide selection are not Umited to those described above.
"Purified" refers to molecules, either polynucleotides or polypeptides that are isolated or separated from their natural environment and are at least 60% free, preferably at least 75% free, and most preferably at least 90% free from other compounds with which they are naturally associated.
A "recombinant nucleic acid" is a sequence that is not naturally occurring or has a sequence that is made by an artificial combination of two or more otherwise separated segments of sequence.
This artificial combination is often accompUshed by chemical synthesis or, more commonly, by the artificial manipulation of isolated segments of nucleic acids , e. g. , by genetic engineering techniques such as those described in Sambrook, supra. The term recombinant includes nucleic acids that have been altered solely by addition, substitution, or deletion of a portion of the nucleic acid. Frequently, a recombinant nucleic acid may include a nucleic acid sequence operably Unked to a promoter sequence.
Such a recombinant nucleic acid may be part of a vector that is used, for example, to transform a cell. Aternatively, such recombinant nucleic acids may be part of a viral vector, e.g., based on a vaccinia virus, that could be use to vaccinate a mammal wherein the recombinant nucleic acid is expressed, inducing a protective immunological response in the mammal.
"Regulatory element" refers to a nucleic acid sequence from nontranslated regions of a gene, and includes enhancers, promoters, introns, and 3' untranslated regions, which interact with host proteins to carry out or regulate transcription or translation.
"Reporter" molecules are chemical or biochemical moieties used for labeUng a nucleic acid, an amino acid, or an antibody. They include radionucUdes; enzymes; fluorescent, chemiluminescent, or chromogenic agents; substrates; cofactors; inhibitors; magnetic particles; and other moieties known in the art. Ai "RNA equivalent," in reference to a DNA sequence, is composed of the same Unear sequence of nucleotides as the reference DNA sequence with the exception that all occurrences of the nitrogenous base thymine are replaced with uracil, and the sugar backbone is composed of ribose instead of deoxyribose. "Sample" is used in its broadest sense. Samples may contain nucleic or amino acids, antibodies, or other materials, and may be derived from any source (e.g., bodily fluids including, but not Umited to, saUva, blood, and urine; chromosome(s), organelles, or membranes isolated from a cell; genomic DNA, RNA, or cDNA in solution or bound to a substrate; and cleared cells or tissues or blots 5 or imprints from such cells or tissues).
"Specific binding" or "specifically binding" refers to the interaction between a protein or peptide and its agonist, antibody, antagonist, or other binding partner. The interaction is dependent upon the presence of a particular structure of the protein, e.g., the antigenic determinant or epitope, recognized by the binding molecule. For example, if an antibody is specific for epitope "A," the o presence of a polypeptide containing epitope A, or the presence of free unlabeled A, in a reaction containing free labeled A and the antibody will reduce the amount of labeled A that binds to the antibody.
"Substitution" refers to the replacement of at least one nucleotide or amino acid by a different nucleotide or amino acid. 5 "Substrate" refers to any suitable rigid or semi-rigid support including, e.g., membranes, filters, chips, sUdes, wafers, fibers, magnetic or nonmagnetic beads, gels, tubing, plates, polymers, microparticles or capillaries. The substrate can have a variety of surface forms, such as wells, trenches, pins, channels and pores, to which polynucleotides or polypeptides are bound.
A "transcript image" refers to the collective pattern of gene expression by a particular tissue or o cell type under given conditions at a given time.
"Transformation" refers to a process by which exogenous DNA enters a recipient cell. Transformation may occur under natural or artificial conditions using various methods well known in the art. Transformation may rely on any known method for the insertion of foreign nucleic acid sequences into a prokaryotic or eukaryotic host cell. The method is selected based on the host cell being 5 transformed.
"Transformants" include stably transformed cells in which the inserted DNA is capable of replication either as an autonomously repUcating plasmid or as part of the host chromosome, as well as cells which transiently express inserted DNA or RNA.
A "transgenic organism," as used herein, is any organism, including but not Umited to animals o and plants, in which one or more of the cells of the organism contains heterologous nucleic acid introduced by way of human intervention, such as by transgenic techniques well known in the art. The nucleic acid is introduced into the cell, directly or indirectly by introduction into a precursor of the cell, by way of deUberate genetic manipulation, such as by microinjection or by infection with a recombinant virus. The term genetic manipulation does not include classical cross-breeding, or in vitro fertiUzation, but rather is directed to the introduction of a recombinant DNA molecule. The transgenic organisms contemplated in accordance with the present invention include bacteria, cyanobacteria, fungi, and plants and animals. The isolated DNA of the present invention can be introduced into the host by methods known in the art, for example infection, transfection, transformation or transconjugatioa Techniques for transferring the DNA of the present invention into such organisms are widely known and provided in references such as Sambrook et al. (1989), supra.
A "variant" of a particular nucleic acid sequence is defined as a nucleic acid sequence having at least 25% sequence identity to the particular nucleic acid sequence over a certain length of one of the nucleic acid sequences using blastn with the "BLAST 2 Sequences" tool Version 2.0.9 (May-07-1999) set at default parameters. Such a pair of nucleic acids may show, for example, at least 30%, at least 50%, at least 60%, at least 70%, at least 80%, at least 90%, at least 95% or even at least 98% or greater sequence identity over a certain defined length. The variant may result in "conservative" amino acid changes which do not affect structural and/or chemical properties. A variant may be described as, for example, an "allelic" (as defined above), "splice," "species," or "polymoφhic" variant. A spUce variant may have significant identity to a reference molecule, but will generally have a greater or lesser number of polynucleotides due to alternate spUcing of exons during mRNA processing. The corresponding polypeptide may possess additional functional domains or lack domains that are present in the reference molecule. Species variants are polynucleotide sequences that vary from one species to another. The resulting polypeptides generally will have significant amino acid identity relative to each other. A polymoφhic variant is a variation in the polynucleotide sequence of a particular gene between individuals of a given species. Polymoφhic variants also may encompass "single nucleotide polymorphisms" (SNPs) in which the polynucleotide sequence varies by one base. The presence of SNPs may be indicative of, for example, a certain population, a disease state, or a propensity for a disease state. In an alternative, variants of the polynucleotides of the present invention may be generated through recombinant methods. One possible method is a DNA shuffling technique such as MOLECULARBREEDING (Maxygen Inc., Santa Clara CA; described in U.S. Patent Number 5,837,458; Chang, C-C et al. (1999) Nat. Biotechnol. 17:793-797; Christians, F.C et al. (1999) Nat. Biotechnol. 17:259-264; and Crameri, A et al. (1996) Nat. Biotechnol. 14:315-319) to alter or improve the biological properties of DITHP, such as its biological or enzymatic activity or its abiUty to bind to other molecules or compounds. DNA shuffling is a process by which a Ubrary of gene variants is produced using PCR-mediated recombination of gene fragments. The Ubrary is then subjected to selection or screening procedures that identify those gene variants with the desired properties. These preferred variants may then be pooled and further subjected to recursive rounds of DNA shuffling and selection/screening. Thus, genetic diversity is created through "artificial" breeding and rapid molecular evolution. For example, fragments of a single gene containing random point mutations may be recombined, screened, and then reshuffled until the desired properties are optimized. Aternatively, fragments of a given gene may be recombined with fragments of homologous genes in the same gene family, either from the same or different species, thereby maximizing the genetic diversity of multiple naturally occurring genes in a directed and controllable manner.
A "variant" of a particular polypeptide sequence is defined as a polypeptide sequence having at least 40% sequence identity to the particular polypeptide sequence over a certain length of one of the polypeptide sequences using blastp with the "BLAST 2 Sequences" tool Version 2.0.9 (May-07- 1999) set at default parameters. Such a pair of polypeptides may show, for example, at least 50%, at least 60%, at least 70%, at least 80%, at least 90%, at least 95%, or at least 98% or greater sequence identity over a certain defined length of one of the polypeptides.
THE INVENTION In a particular embodiment, cDNA sequences derived from human tissues and cell Unes were aUgned based on nucleotide sequence identity and assembled into "consensus" or "template" sequences which are designated by the template identification numbers (template IDs) in column 2 of Table 1. The sequence identification numbers (SEQ ID NO:s) corresponding to the template IDs are shown in column 1. The template sequences have similarity to GenBank sequences, or "hits," as designated by the GI Numbers in column 3. The statistical probability of each GenBank hit is indicated by a probabiUty score in column 4, and the functional annotation corresponding to each GenBank hit is Usted in column 5.
The invention incoφorates the nucleic acid sequences of these templates as disclosed in the Sequence Listing and the use of these sequences in the diagnosis and treatment of disease states characterized by defects in human molecules. The invention further utiUzes these sequences in hybridization and ampUfication technologies, and in particular, in technologies which assess gene expression patterns correlated with specific cells or tissues and their responses in vivo or in vitro to pharmaceutical agents, toxins, and other treatments. In this manner, the sequences of the present invention are used to develop a transcript image for a particular cell or tissue.
Derivation of Nucleic Acid Sequences cDNA was isolated from Ubraries constructed using RNA derived from normal and diseased human tissues and cell Unes. The human tissues and cell Unes used for cDNA Ubrary construction were selected from a broad range of sources to provide a diverse population of cDNA representative of gene transcription throughout the human body. Descriptions of the human tissues and cell Unes used for cDNA Ubrary construction are provided in the LEFESEQ database (Incyte Genomics, Inc. (Incyte), Palo Ato CA). Human tissues were broadly selected from, for example, cardiovascular, dermatologic, endocrine, gastrointestinal, hematopoietic/immune system, musculoskeletal, neural, reproductive, and urologic sources.
Cell Unes used for cDNA library construction were derived from, for example, leukemic cells, teratocarcinomas, neuroepitheUomas, cervical carcinoma, lung fibroblasts, and endotheUal cells. Such cell lines include, for example, THP-1, Jurkat, HUVEC, hNT2, WI38, HeLa, and other cell Unes commonly used and available from pubUc depositories (American Type Culture Collection, Manassas VA). Prior to mRNA isolation, cell Unes were untreated, treated with a pharmaceutical agent such as 5 -aza-2'-deoxycytidine, treated with an activating agent such as Upopolysaccharide in the case of leukocytic cell Unes, or, in the case of endotheUal cell lines, subjected to shear stress.
Sequencing of the cDNA Methods for DNA sequencing are well known in the art. Conventional enzymatic methods employ the Klenow fragment of DNA polymerase I, SEQUENASE DNA polymerase (U.S. Biochemical Coφoration, Cleveland OH), Taq polymerase (PE Biosystems, Foster City CA), thermostable T7 polymerase (Amersham Pharmacia Biotech, Inc. (Amersham Pharmacia Biotech), Piscataway NJ), or combinations of polymerases and proofreading exonucleases such as those found in the ELONGASE ampUfication system (Life Technologies Inc. (Life Technologies), Gaithersburg MD), to extend the nucleic acid sequence from an oUgonucleotide primer annealed to the DNA template of interest. Methods have been developed for the use of both single-stranded and double-stranded templates. Chain termination reaction products may be electrophoresed on urea-polyacrylamide gels and detected either by autoradiography (for radioisotope-labeled nucleotides) or by fluorescence (for fluorophore-labeled nucleotides). Automated methods for mechanized reaction preparation, sequencing, and analysis using fluorescence detection methods have been developed. Machines used to prepare cDNA for sequencing can include the MICROLAB 2200 liquid transfer system (Hamilton Company (Hamilton), Reno NV), Peltier thermal cycler (PTC200; MJ Research, Inc. (MJ Research), Watertown MA), and ABI CATALYST 800 thermal cycler (PE Biosystems). Sequencing can be carried out using, for example, the ABI 373 or 377 (PE Biosystems) or MEGABACE 1000 (Molecular Dynamics, Inc.
(Molecular Dynamics), Sunnyvale CA) DNA sequencing systems, or other automated and manual sequencing systems well known in the art.
The nucleotide sequences of the Sequence Listing have been prepared by current, state-of-the- art, automated methods and, as such, may contain occasional sequencing errors or unidentified nucleotides. Such unidentified nucleotides are designated by an N. These infrequent unidentified bases do not represent a hindrance to practicing the invention for those skilled in the art. Several methods employing standard recombinant techniques may be used to correct errors and complete the missing sequence information. (See, e.g., those described in Ausubel, F.M. et al. (1997) Short Protocols in 5 Molecular Biology, John Wiley & Sons, New York NY; and Sambrook, J. et al. (1989) Molecular Cloning, A Laboratory Manual, Cold Spring Harbor Press, Plainview NY.)
Asembly of cDNA Sequences
Human polynucleotide sequences may be assembled using programs or algorithms well known 0 in the art. Sequences to be assembled are related, wholly or in part, and may be derived from a single or many different transcripts. Asembly of the sequences can be performed using such programs as PHRAP (Phils Revised Asembly Program) and the GEL VIEW fragment assembly system (GCG), or other methods known in the art.
Aternatively, cDNA sequences are used as "component" sequences that are assembled into 5 "template" or "consensus" sequences as follows. Sequence chromatograms are processed, verified, and quaUty scores are obtained using PHRED. Raw sequences are edited using an editing pathway known as Block 1 (See, e.g., theLIFESEQ Asembled User Guide, Incyte Genomics, Palo Ato, CA). A series of BLAST comparisons is performed and low-information segments and repetitive elements (e.g., dinucleotide repeats, Au repeats, etc.) are replaced by "n's", or masked, to prevent spurious matches. o Mitochondrial and ribosomal RNA sequences are also removed. The processed sequences are then loaded into a relational database management system (RDMS) which assigns edited sequences to existing templates, if available. When additional sequences are added into the RDMS, a process is initiated which modifies existing templates or creates new templates from works in progress (i.e., nonfinal assembled sequences) containing queued sequences or the sequences themselves. After the new 5 sequences have been assigned to templates, the templates can be merged into bins. If multiple templates exist in one bin, the bin can be spUt and the templates reannotated.
Once gene bins have been generated based upon sequence alignments, bins are "clone joined" based upon clone information. Clone joining occurs when the 5' sequence of one clone is present in one bin and the 3' sequence from the same clone is present in a different bin, indicating that the two bins o should be merged into a single bin. Only bins which share at least two different clones are merged.
A resultant template sequence may contain either a partial or a full length open reading frame, or all or part of a genetic regulatory element. This variation is due in part to the fact that the full length cDNA of many genes are several hundred, and sometimes several thousand, bases in length. With current technology, cDNA comprising the coding regions of large genes cannot be cloned because of vector Umitations, incomplete reverse transcription of the mRNA, or incomplete "second strand" synthesis. Template sequences may be extended to include additional contiguous sequences derived from the parent RNA transcript using a variety of methods known to those of skill in the art. Extension may thus be used to achieve the full length coding sequence of a gene.
5
Analysis of the cDNA Sequences
The cDNA sequences are analyzed using a variety of programs and algorithms which are well known in the art. (See, e.g., Ausubel, 1997, supra. Chapter 7.7; Meyers, R.A. (Ed.) (1995) Molecular Biology and Biotechnology, Wiley VCH, New York NY, pp. 856-853; and Table 6.) These analyses o comprise both reading frame determinations, e.g. , based on triplet codon periodicity for particular organisms (Fickett, J.W. (1982) Nucleic Acids Res. 10:5303-5318); analyses of potential start and stop codons; and homology searches.
Computer programs known to those of skill in the art for performing computer-assisted searches for amino acid and nucleic acid sequence similarity, include, for example, Basic Local 5 AUgnment Search Tool (BLAST; Atschul, S.F. (1993) J. Mol. Evol. 36:290-300; Atschul, S.F. et al. (1990) J. Mol. Biol. 215:403-410). BLAST is especially useful in determining exact matches and comparing two sequence fragments of arbitrary but equal lengths, whose aUgnment is locally maximal and for which the alignment score meets or exceeds a threshold or cutoff score set by the user (KarUn, S. et al. (1988) Proc. Natl. Acad. Sci. USA 85:841-845). Using an appropriate search tool (e.g., o BLAST or HMM), GenBank, SwissProt, BLOCKS, PFAM and other databases may be searched for sequences containing regions of homology to a query dithp or DITHP of the present invention.
Other approaches to the identification, assembly, storage, and display of nucleotide and polypeptide sequences are provided in "Relational Database for Storing Biomolecule Information," U.S.S.N. 08/947,845, filed October 9, 1997; "Project-Based Full-Length Biomolecular Sequence 5 Database," U.S.S.N. 08/811,758, filed March 6, 1997; and "Relational Database and System for
Storing Information Relating to Biomolecular Sequences," U.S.S.N. 09/034,807, filed March 4, 1998, all of which are incoφorated by reference herein in their entirety.
Protein hierarchies can be assigned to the putative encoded polypeptide based on, e.g., motif, BLAST, or biological analysis. Methods for assigning these hierarchies are described, for example, in o "Database System Employing Protein Function Hierarchies for Viewing Biomolecular Sequence Data,"
U.S.S.N. 08/812,290, filed March 6, 1997, incoφorated herein by reference.
Identification of Human Diagnostic and Therapeutic Molecules Encoded by dithp The identities of the DITHP encoded by the dithp of the present invention were obtained by analysis of the assembled cDNA sequences. SEQ ID NO:l, SEQ ID NO:2, SEQ ID NO:3, SEQ ID NO:4, SEQ ID NO:5, SEQ ID NO:6, SEQ ID NO:7, and SEQ ID NO:8 encode, for example, human enzyme molecules. SEQ ID NO:9 encodes, for example, an extracellular information transmission molecule. SEQ ID NO: 10 and SEQ ID NO: 11 encode, for example, receptor molecules. SEQ ID NO:12, SEQ ID NO:13, SEQ ID NO:14, SEQ ID NO:15, SEQ ID NO:16, SEQ ID NO:17, and SEQ ID NO:18 encode, for example, intracellular signaling molecules. SEQ ID NO:19, SEQ ID NO:20, SEQ ID NO:21, SEQ ID NO:22, SEQ ID NO:23, SEQ ID NO:24, SEQ ID NO:25, SEQ ID NO:26, SEQ ID NO:27, SEQ ID NO:28, SEQ ID NO:29, SEQ ID NO:30, SEQ ID NO:31, SEQ ID NO:32, and SEQ ID NO:33 encode, for example, ttanscription factor molecules. SEQ ID NO:34 encodes, for example, a protein modification and maintenance molecule. SEQ ID NO:35 and SEQ ID NO:36 encode, for example, nucleic acid synthesis and modification molecules. SEQ ID NO:37 encodes, for example, an antigen recognition molecule. SEQ ID NO:38 and SEQ ID NO:39 encode, for example, secreted/extracellular matrix molecules. SEQ ID NO:40, SEQ ID NO:41, SEQ ID NO:42, SEQ ID NO:43, SEQ ID NO:44, and SEQ ID NO:45 encode, for example, cytoskeletal molecules. SEQ ID NO:46, SEQ ID NO:47, and SEQ ID NO:48 encode, for example, cell membrane molecules. SEQ ID NO:49, SEQ ID NO:50, SEQ ID NO:51, SEQ ID NO:52, and SEQ ID NO:53 encode, for example, ribosomal molecules. SEQ ID NO:54, SEQ ID NO:55, SEQ ID NO:56, SEQ ID NO:57, SEQ ID NO:58, SEQ ID NO:59, SEQ ID NO:60, SEQ ID NO:61, SEQ ID NO:62, and SEQ ID NO:63 encode, for example, organelle associated molecules. SEQ ID NO:64, SEQ ID NO:65, SEQ ID NO:66, SEQ ID NO:67, and SEQ ID NO:68 encode, for example, biochemical pathway molecules. SEQ ID NO:69, SEQ ID NO:70, and SEQ ID NO:71 encode, for example, molecules associated with growth and development.
Sequences of Human Diagnostic and Therapeutic Molecules
The dithp of the present invention may be used for a variety of diagnostic and therapeutic puφoses. For example, a dithp may be used to diagnose a particular condition, disease, or disorder associated with human molecules. Such conditions, diseases, and disorders include, but are not Umited to, a cell proliferative disorder, such as actinic keratosis, arteriosclerosis, atherosclerosis, bursitis, cirrhosis, hepatitis, mixed connective tissue disease (MCTD), myelofibrosis, paroxysmal nocturnal hemoglobinuria, polycythemia vera, psoriasis, primary thrombocythemia, and cancers including adenocarcinoma, leukemia, lymphoma, melanoma, myeloma, sarcoma, teratocarcinoma, and, in particular, a cancer of the adrenal gland, bladder, bone, bone marrow, brain, breast, cervix, gall bladder, ganglia, gastrointestinal tract, heart, kidney, liver, lung, muscle, ovary, pancreas, parathyroid, penis, prostate, salivary glands, skin, spleen, testis, thymus, thyroid, and uterus; an autoimmune/inflammatory disorder, such as inflammation, actinic keratosis, acquired immunodeficiency syndrome (ADS), Addison's disease, adult respiratory distress syndrome, allergies, ankylosing spondylitis, amyloidosis, anemia, arteriosclerosis, asthma, atherosclerosis, 5 autoimmune hemolytic anemia, autoimmune thyroiditis, bronchitis, bursitis, cholecystitis, cirrhosis, contact dermatitis, Crohn's disease, atopic dermatitis, dermatomyositis, diabetes melUtus, emphysema, erythroblastosis fetalis, erythema nodosum, atrophic gastritis, glomerulonephritis, Goodpasture's syndrome, gout, Graves' disease, Hashimoto's thyroiditis, paroxysmal nocturnal hemoglobinuria, hepatitis, hypereosinophiUa, irritable bowel syndrome, episodic lymphopenia with 0 lymphocytotoxins, mixed connective tissue disease (MCTD), multiple sclerosis, myasthenia gravis, myocardial or pericardial inflammation, myelofibrosis, osteoarthritis, osteoporosis, pancreatitis, polycythemia vera, polymyositis, psoriasis, Reiter's syndrome, rheumatoid arthritis, scleroderma, Sjogren's syndrome, systemic anaphylaxis, systemic lupus erythematosus, systemic sclerosis, primary thrombocythemia, thrombocytopenic puφura, ulcerative colitis, uveitis, Werner syndrome, 5 complications of cancer, hemodialysis, and extracoφoreal circulation, trauma, and hematopoietic cancer including lymphoma, leukemia, and myeloma; an infection caused by a viral agent classified as adenovirus, arenavirus, bunyavirus, calicivirus, coronavirus, filovirus, hepadnavirus, heφesvirus, flavivirus, orthomyxovirus, parvovirus, papovavirus, paramyxovirus, picornavirus, poxvirus, reovirus, retrovirus, rhabdovirus, or togavirus; an infection caused by a bacterial agent classified as 0 pneumococcus, staphylococcus, streptococcus, bacillus, corynebacterium, clostridium, meningococcus, gonococcus, listeria, moraxella, kingella, haemophilus, legionella, bordetella, gram- negative enterobacterium including shigella, salmonella, or campylobacter, pseudomonas, vibrio, brucella, francisella, yersinia, bartonella, norcardium, actinomyces, mycobacterium, spirochaetale, rickettsia, chlamydia, or mycoplasma; an infection caused by a fungal agent classified as aspergillus, 5 blastomyces, dermatophytes, cryptococcus, coccidioides, malasezzia, histoplasma, or other mycosis- causing fungal agent; and an infection caused by a parasite classified as plasmodium or malaria- causing, parasitic entamoeba, leishmania, trypanosoma, toxoplasma, pneumocystis carinii, intestinal protozoa such as giardia, trichomonas, tissue nematode such as trichinella, intestinal nematode such as ascaris, lymphatic filarial nematode, trematode such as schistosoma, and cesttode such as o tapeworm; a developmental disorder such as renal tubular acidosis, anemia, Cushing's syndrome, achondroplastic dwarfism, Duchenne and Becker muscular dystrophy, epilepsy, gonadal dysgenesis, WAGR syndrome (Wilms' tumor, aniridia, genitourinary abnormalities, and mental retardation), Smith-Magenis syndrome, myelodysplastic syndrome, hereditary mucoepithelial dysplasia, hereditary keratodermas, hereditary neuropathies such as Charcot-Marie-Tooth disease and neurofibromatosis, 5 hypothyroidism, hydrocephalus, seizure disorders such as Syndenham's chorea and cerebral palsy, spina bifida, anencephaly, craniorachischisis, congenital glaucoma, cataract, and sensorineural hearing loss; an endocrine disorder such as a disorder of the hypothalamus and/or pituitary resulting from lesions such as a primary brain tumor, adenoma, infarction associated with pregnancy, hypophysectomy, aneurysm, vascular malformation, thrombosis, infection, immunological disorder, and 5 complication due to head trauma; a disorder associated with hypopituitarism including hypogonadism, Sheehan syndrome, diabetes insipidus, Kallman's disease, Hand-Schuller-Christian disease, Letterer- Siwe disease, sarcoidosis, empty sella syndrome, and dwarfism; a disorder associated with hypeφituitarism including acromegaly, giantism, and syndrome of inappropriate antidiuretic hormone (ADH) secretion (SIADH) often caused by benign adenoma; a disorder associated with hypothyroidism 0 including goiter, myxedema, acute thyroiditis associated with bacterial infection, subacute thyroiditis associated with viral infection, autoimmune thyroiditis (Hashimoto's disease), and cretinism; a disorder associated with hyperthyroidism including thyrotoxicosis and its various forms, Grave's disease, pretibial myxedema, toxic multinodular goiter, thyroid carcinoma, and Plummer's disease; a disorder associated with hypeφarathyroidism including Conn disease (chronic hypercalemia); a pancreatic 5 disorder such as Type I or Type II diabetes melUtus and associated compUcations; a disorder associated with the adrenals such as hypeφlasia, carcinoma, or adenoma of the adrenal cortex, hypertension associated with alkalosis, amyloidosis, hypokalemia, Cushing's disease, Liddle's syndrome, and Anold-Healy-Gordon syndrome, pheochromocytoma tumors, and Addison's disease; a disorder associated with gonadal steroid hormones such as: in women, abnormal prolactin production, o infertiUty, endometriosis, perturbation of the menstrual cycle, polycystic ovarian disease, hypeφrolactinemia, isolated gonadotropin deficiency, amenorrhea, galactorrhea, hermaphroditism, hirsutism and viriUzation, breast cancer, and, in post-menopausal women, osteoporosis; and, in men, Leydig cell deficiency, male cUmacteric phase, and germinal cell aplasia, a hypergonadal disorder associated with Leydig cell tumors, androgen resistance associated with absence of androgen receptors, 5 syndrome of 5 α-reductase, and gynecomastia; a metabolic disorder such as Addison's disease, cerebrotendinous xanthomatosis, congenital adrenal hypeφlasia, coumarin resistance, cystic fibrosis, diabetes, fatty hepatocirrhosis, fructose- 1,6-diphosphatase deficiency, galactosemia, goiter, glucagonoma, glycogen storage diseases, hereditary fructose intolerance, hyperadrenalism, hypoadrenalism, hypeφarathyroidism, hypoparathyroidism, hypercholesterolemia, hyperthyroidism, o hypoglycemia, hypothyroidism, hyperlipidemia, hyperlipemia, lipid myopathies, lipodystrophies, lysosomal storage diseases, mannosidosis, neuraminidase deficiency, obesity, pentosuria phenylketonuria, pseudovitamin D-deficiency rickets; disorders of carbohydrate metabolism such as congenital type II dyserythropoietic anemia, diabetes, insulin-dependent diabetes melUtus, non-insulin-dependent diabetes melUtus, fructose- 1,6-diphosphatase deficiency, galactosemia, glucagonoma, hereditary fructose intolerance, hypoglycemia, mannosidosis, neuraminidase deficiency, obesity, galactose epimerase deficiency, glycogen storage diseases, lysosomal storage diseases, fructosuria, pentosuria, and inherited abnormalities of pyruvate metabolism; disorders of lipid metabolism such as fatty liver, cholestasis, primary biliary cirrhosis, carnitine deficiency, carnitine palmitoylttansferase deficiency, myoadenylate deaminase deficiency, hypertriglyceridemia, lipid storage disorders such Fabry's disease, Gaucher's disease, Niemann-Pick' s disease, metachromatic leukodysfrophy, adrenoleukodystrophy, GM^ gangUosidosis, and ceroid lipofuscinosis, abetalipoproteinemia, Tangier disease, hyperlipoproteinemia, diabetes melUtus, lipodystrophy, lipomatoses, acute panniculitis, disseminated fat necrosis, adiposis dolorosa, lipoid adrenal hypeφlasia, minimal change disease, lipomas, atherosclerosis, hypercholesterolemia, hypercholesterolemia with hypertriglyceridemia, primary hypoalphalipoproteinemia, hypothyroidism, renal disease, liver disease, lecithinxholesterol acyltransferase deficiency, cerebrotendinous xanthomatosis, sitosterolemia, hypocholesterolemia, Tay-Sachs disease, Sandhoff s disease, hyperUpidemia, hyperlipemia, lipid myopathies, and obesity; and disorders of copper metabolism such as Menke's disease, Wilson's disease, and Ehlers-Danlos syndrome type IX; a neurological disorder such as epilepsy, ischemic cerebrovascular disease, stroke, cerebral neoplasms, Azheimer' s disease, Pick's disease, Huntington's disease, dementia, Parkinson's disease and other extrapyramidal disorders, amyotrophic lateral sclerosis and other motor neuron disorders, progressive neural muscular atrophy, retinitis pigmentosa, hereditary ataxias, multiple sclerosis and other demyeUnating diseases, bacterial and viral meningitis, brain abscess, subdural empyema, epidural abscess, suppurative intracranial thrombophlebitis, myeUtis and radicuUtis, viral central nervous system disease, prion diseases including kuru, Creutzfeldt-Jakob disease, and Gerstmann-Straussler-Scheinker syndrome, fatal famiUal insomnia, nutritional and metabolic diseases of the nervous system, neurofibromatosis, tuberous sclerosis, cerebelloretinal hemangioblastomatosis, encephalotrigeminal syndrome, mental retardation and other developmental disorder of the central nervous system, cerebral palsy, a neuroskeletal disorder, an autonomic nervous system disorder, a cranial nerve disorder, a spinal cord disease, muscular dystrophy and other neuromuscular disorder, a peripheral nervous system disorder, dermatomyositis and polymyositis, inherited, metabolic, endocrine, and toxic myopathy, myasthenia gravis, periodic paralysis, a mental disorder including mood, anxiety, and schizophrenic disorders, seasonal affective disorder (SAD), akathesia, amnesia, catatonia, diabetic neuropathy, tardive dyskinesia, dystonias, paranoid psychoses, postheφetic neuralgia, and Tourette's disorder; a gastrointestinal disorder including ulcerative colitis, gastric and duodenal ulcers, cystinuria, dibasicaminoaciduria, hypercystinuria, lysinuria, hartnup disease, tryptophan malabsoφtion, methionine malabsoφtion, histidinuria, iminoglycinuria, dicarboxylicaminoaciduria, cystinosis, renal glycosuria, hypouricemia, familial hypophophatemic rickets, congenital chloridorrhea, distal renal tubular acidosis, Menkes' disease, Wilson's disease, lethal diarrhea, juvenile pernicious anemia, ate malabsoφtion, adrenoleukodystrophy, hereditary myoglobinuria, and Zellweger syndrome; a transport disorder such as akinesia, amyotrophic lateral sclerosis, ataxia telangiectasia, cystic fibrosis, Becker's muscular dystrophy, Bell's palsy, Charcot-Marie Tooth disease, diabetes mellitus, diabetes insipidus, diabetic neuropathy, Duchenne muscular dystrophy, hyperkalemic periodic paralysis, normokalemic periodic paralysis, Parkinson's disease, maUgnant hyperthermia, multidrug resistance, myasthenia gravis, myotonic dystrophy, catatonia, tardive dyskinesia, dystonias, peripheral neuropathy, cerebral neoplasms, prostate cancer, cardiac disorders associated with transport, e.g., angina, bradyarrythmia, tachyarrythmia, hypertension, Long QT syndrome, myocarditis, cardiomyopathy, nemaline myopathy, centronuclear myopathy, lipid myopathy, mitochondrial myopathy, thyrotoxic myopathy, ethanol myopathy, dermatomyositis, inclusion body myositis, infectious myositis, and polymyositis, neurological disorders associated with transport, e.g., Azheimer' s disease, amnesia, bipolar disorder, dementia, depression, epilepsy, Tourette's disorder, paranoid psychoses, and schizophrenia, and other disorders associated with transport, e.g., neurofibromatosis, postheφetic neuralgia, trigeminal neuropathy, sarcoidosis, sickle cell anemia, cataracts, infertility, pulmonary artery stenosis, sensorineural autosomal deafness, hyperglycemia, hypoglycemia, Grave's disease, goiter, glucose-galactose malabsoφtion syndrome, hypercholesterolemia, Cushing's disease, and Addison's disease; and a connective tissue disorder such as osteogenesis imperfecta, Ehlers-Danlos syndrome, chondrodysplasias, Marfan syndrome, Aport syndrome, familial aortic aneurysm, achondroplasia, mucopolysaccharidoses, osteoporosis, osteopetrosis, Paget's disease, rickets, osteomalacia, hypeφarathyroidism, renal osteodystrophy, osteonecrosis, osteomyelitis, osteoma, osteoid osteoma, osteoblastoma, osteosarcoma, osteochondroma, chondroma, chondroblastoma, chondromyxoid fibroma, chondrosarcoma, fibrous cortical defect, nonossifying fibroma, fibrous dysplasia, fibrosarcoma, malignant fibrous histiocytoma, Ewing's sarcoma, primitive neuroectodermal tumor, giant cell tumor, osteoarthritis, rheumatoid arthritis, ankylosing spondyloarthritis, Reiter's syndrome, psoriatic arthritis, enteropathic arthritis, infectious arthritis, gout, gouty arthritis, calcium pyrophosphate crystal deposition disease, ganglion, synovial cyst, villonodular synovitis, systemic sclerosis, Dupuytren's contracture, hepatic fibrosis, lupus erythematosus, mixed connective tissue disease, epidermolysis bullosa simplex, buUous congenital ichthyosiform erythroderma (epidermolytic hyperkeratosis), non-epidermolytic and epidermolytic palmoplantar keratoderma, ichthyosis bullosa of Siemens, pachyonychia congenita, and white sponge nevus. The dithp can be used to detect the presence of, or to quantify the amount of, a dithp-related polynucleotide in a sample. This information is then compared to information obtained from appropriate reference samples, and a diagnosis is estabUshed. Aternatively, a polynucleotide complementary to a given dithp can inhibit or inactivate a therapeutically relevant gene related to the dithp.
A alvsis of dithp Expression Patterns
The expression of dithp may be routinely assessed by hybridization-based methods to determine, for example, the tissue-specificity, disease-specificity, or developmental stage-specificity of dithp expression. For example, the level of expression of dithp may be compared among different cell types or tissues, among diseased and normal cell types or tissues, among cell types or tissues at different developmental stages, or among cell types or tissues undergoing various treatments. This type of analysis is useful, for example, to assess the relative levels of dithp expression in fully or partially differentiated cells or tissues, to determine if changes in dithp expression levels are correlated with the development or progression of specific disease states, and to assess the response of a cell or tissue to a specific therapy, for example, in pharmacological or toxicological studies. Methods for the analysis of dithp expression are based on hybridization and ampUfication technologies and include membrane- based procedures such as northern blot analysis, high-throughput procedures that utitize, for example, microarrays, and PCR-based procedures.
Hybridization and Genetic Aialvsis The dithp, their fragments, or complementary sequences, may be used to identify the presence of and or to determine the degree of similarity between two (or more) nucleic acid sequences. The dithp may be hybridized to naturally occurring or recombinant nucleic acid sequences under appropriately selected temperatures and salt concentrations. Hybridization with a probe based on the nucleic acid sequence of at least one of the dithp allows for the detection of nucleic acid sequences, including genomic sequences, which are identical or related to the dithp of the Sequence Listing. Probes may be selected from non-conserved or unique regions of at least one of the polynucleotides of SEQ ID NO:l- 71 and tested for their abiUty to identify or ampUfy the target nucleic acid sequence using standard protocols.
Polynucleotide sequences that are capable of hybridizing, in particular, to those shown in SEQ ID NO: 1-71 and fragments thereof, can be identified using various conditions of stringency. (See, e.g. , Wahl, G.M. and S.L. Berger (1987) Methods Enzymol. 152:399-407; Kimmel, AR. (1987) Methods Enzymol. 152:507-511.) Hybridization conditions are discussed in "Definitions."
A probe for use in Southern or northern hybridization may be derived from a fragment of a dithp sequence, or its complement, that is up to several hundred nucleotides in length and is either single-stranded or double-stranded. Such probes may be hybridized in solution to biological materials such as plasmids, bacterial, yeast, or human artificial chromosomes, cleared or sectioned tissues, or to artificial substrates containing dithp. Microarrays are particularly suitable for identifying the presence of and detecting the level of expression for multiple genes of interest by examining gene expression 5 correlated with, e.g., various stages of development, treatment with a drug or compound, or disease progression. An array analogous to a dot or slot blot may be used to arrange and tink polynucleotides to the surface of a substrate using one or more of the following: mechanical (vacuum), chemical, thermal, or UV bonding procedures. Such an array may contain any number of dithp and may be produced by hand or by using available devices, materials, and machines. 0 Microarrays may be prepared, used, and analyzed using methods known in the art. (See, e.g.,
Brennan, T.M. et al. (1995) U.S. Patent No. 5,474,796; Schena, M. et al. (1996) Proc. Natl. Acad. Sci. USA 93:10614-10619; Baldeschweiler et al. (1995) PCT application W095/251116; Shalon, D. et al. (1995) PCT appUcation WO95/35505; Heller, R.A. et al. (1997) Proc. Natl. Acad. Sci. USA 94:2150- 2155; and Heller, MJ. et al. (1997) U.S. Patent No. 5,605,662.) 5 Probes may be labeled by either PCR or enzymatic techniques using a variety of commercially available reporter molecules. For example, commercial kits are available for radioactive and chemiluminescent labeling (Amersham Pharmacia Biotech) and for alkaline phosphatase labeUng (Life Technologies). Aternatively, dithp may be cloned into commercially available vectors for the production of RNA probes. Such probes may be transcribed in the presence of at least one labeled o nucleotide (e.g. , 32P-ATP, Amersham Pharmacia Biotech).
Additionally the polynucleotides of SEQ ID NO: 1-71 or suitable fragments thereof can be used to isolate full length cDNA sequences utihzing hybridization and/or ampUfication procedures well known in the art, e.g., cDNA Ubrary screening, PCR ampUfication, etc. The molecular cloning of such full length cDNA sequences may employ the method of cDNA library screening with probes using the 5 hybridization, stringency, washing, and probing strategies described above and in Ausubel, supra.
Chapters 3, 5, and 6. These procedures may also be employed with genomic Ubraries to isolate genomic sequences of dithp in order to analyze, e.g., regulatory elements.
Genetic Mapping o Gene identification and mapping are important in the investigation and treatment of almost all conditions, diseases, and disorders. Cancer, cardiovascular disease, Azheimer's disease, arthritis, diabetes, and mental illnesses are of particular interest. Each of these conditions is more complex than the single gene defects of sickle cell anemia or cystic fibrosis, with select groups of genes being predictive of predisposition for a particular condition, disease, or disorder. For example, cardiovascular disease may result from malfunctioning receptor molecules that fail to clear cholesterol from the bloodstream, and diabetes may result when a particular individual's immune system is activated by an infection and attacks the insuUn-producing cells of the pancreas. In some studies, Azheimer 's disease has been Unked to a gene on chromosome 21 ; other studies predict a different gene 5 and location. Mapping of disease genes is a complex and reiterative process and generally proceeds from genetic tinkage analysis to physical mapping.
As a condition is noted among members of a family, a genetic tinkage map traces parts of chromosomes that are inherited in the same pattern as the conditioa Statistics Unk the inheritance of particular conditions to particular regions of chromosomes, as defined by RFLP or other markers. 0 (See, for example, Lander, E. S. and Botstein, D. (1986) Proc. Natl. Acad. Sci. USA 83:7353-7357.) Occasionally, genetic markers and their locations are known from previous studies. More often, however, the markers are simply stretches of DNA that differ among individuals. Examples of genetic tinkage maps can be found in various scientific journals or at the Onhne Mendelian Inheritance in Man (OMIM) World Wide Web site. 5 In another embodiment of the invention, dithp sequences may be used to generate hybridization probes useful in chromosomal mapping of naturally occurring genomic sequences. Either coding or noncoding sequences of dithp may be used, and in some instances, noncoding sequences may be preferable over coding sequences. For example, conservation of a dithp coding sequence among members of a multi-gene family may potentially cause undesired cross hybridization during o chromosomal mapping. The sequences may be mapped to a particular chromosome, to a specific region of a chromosome, or to artificial chromosome constructions, e.g., human artificial chromosomes (HACs), yeast artificial chromosomes (YACs), bacterial artificial chromosomes (BACs), bacterial PI constructions, or single chromosome cDNA Ubraries. (See, e.g., Harrington, J.J. et al. (1997) Nat. Genet. 15:345-355; Price, CM. (1993) Blood Rev. 7:127-134; and Trask, B J. (1991) Trends Genet. 5 7:149-154.)
Fluorescent in situ hybridization (FISH) may be correlated with other physical chromosome mapping techniques and genetic map data. (See, e.g., Meyers, supra, pp. 965-968.) Correlation between the location of dithp on a physical chromosomal map and a specific disorder, or a predisposition to a specific disorder, may help define the region of DNA associated with that disorder. o The dithp sequences may also be used to detect polymoφhisms that are genetically Unked to the inheritance of a particular condition, disease, or disorder.
In situ hybridization of chromosomal preparations and genetic mapping techniques, such as Unkage analysis using established chromosomal markers, may be used for extending existing genetic maps. Often the placement of a gene on the chromosome of another mammalian species, such as mouse, may reveal associated markers even if the number or arm of the corresponding human chromosome is not known. These new marker sequences can be mapped to human chromosomes and may provide valuable information to investigators searching for disease genes using positional cloning or other gene discovery techniques. Once a disease or syndrome has been crudely correlated by genetic 5 tinkage with a particular genomic region, e.g., ataxia-telangiectasia to 1 lq22-23, any sequences mapping to that area may represent associated or regulatory genes for further investigation. (See, e.g., Gatti, R.A. et al. (1988) Nature 336:577-580.) The nucleotide sequences of the subject invention may also be used to detect differences in chromosomal architecture due to translocation, inversion, etc., among normal, carrier, or affected individuals. o Once a disease-associated gene is mapped to a chromosomal region, the gene must be cloned in order to identify mutations or other alterations (e.g., ttanslocations or inversions) that may be correlated with disease. This process requires a physical map of the chromosomal region containing the disease- gene of interest along with associated markers. A physical map is necessary for determining the nucleotide sequence of and order of marker genes on a particular chromosomal region. Physical 5 mapping techniques are well known in the art and require the generation of overlapping sets of cloned DNA fragments from a particular organelle, chromosome, or genome. These clones are analyzed to reconstruct and catalog their order. Once the position of a marker is determined, the DNA from that region is obtained by consulting the catalog and selecting clones from that region. The gene of interest is located through positional cloning techniques using hybridization or similar methods. 0
Diagnostic Uses
The dithp of the present invention may be used to design probes useful in diagnostic assays. Such assays, well known to those skilled in the art, may be used to detect or confirm conditions, disorders, or diseases associated with abnormal levels of dithp expression. Labeled probes developed 5 from dithp sequences are added to a sample under hybridizing conditions of desired sttingency. In some instances, dithp, or fragments or oUgonucleotides derived from dithp, may be used as primers in ampUfication steps prior to hybridization. The amount of hybridization complex formed is quantified and compared with standards for that cell or tissue. If dithp expression varies significantly from the standard, the assay indicates the presence of the condition, disorder, or disease. QuaUtative or o quantitative diagnostic methods may include northern, dot blot, or other membrane or dip-stick based technologies or multiple-sample format technologies such as PCR, enzyme-Unked immunosorbent assay (ELISA)-Uke, pin, or chip-based assays.
The probes described above may also be used to monitor the progress of conditions, disorders, or diseases associated with abnormal levels of dithp expression, or to evaluate the efficacy of a particular therapeutic treatment. The candidate probe may be identified from the dithp that are specific to a given human tissue and have not been observed in GenBank or other genome databases. Such a probe may be used in animal studies, precUnical tests, cUnical trials, or in monitoring the treatment of an individual patient. In a typical process, standard expression is estabUshed by methods well known in 5 the art for use as a basis of comparison, samples from patients affected by the disorder or disease are combined with the probe to evaluate any deviation from the standard profile, and a therapeutic agent is administered and effects are monitored to generate a treatment profile. Efficacy is evaluated by determining whether the expression progresses toward or returns to the standard normal pattern. Treatment profiles may be generated over a period of several days or several months. Statistical o methods well known to those skilled in the art may be use to determine the significance of such therapeutic agents.
The polynucleotides are also useful for identifying individuals from minute biological samples, for example, by matching the RFLP pattern of a sample's DNA to that of an individual's DNA The polynucleotides of the present invention can also be used to determine the actual base-by-base DNA 5 sequence of selected portions of an individual's genome. These sequences can be used to prepare PCR primers for ampUfying and isolating such selected DNA, which can then be sequenced. Using this technique, an individual can be identified through a unique set of DNA sequences. Once a unique ID database is established for an individual, positive identification of that individual can be made from extremely small tissue samples. 0 In a particular aspect, oUgonucleotide primers derived from the dithp of the invention may be used to detect single nucleotide polymoφhisms (SNPs). SNPs are substitutions, insertions and deletions that are a frequent cause of inherited or acquired genetic disease in humans. Methods of SNP detection include, but are not Umited to, single-stranded conformation polymorphism (SSCP) and fluorescent SSCP (fSSCP) methods. In SSCP, oligonucleotide primers derived from dithp are used to 5 ampUfy DNA using the polymerase chain reaction (PCR). The DNA may be derived, for example, from diseased or normal tissue, biopsy samples, bodily fluids, and the like. SNPs in the DNA cause differences in the secondary and tertiary structures of PCR products in single-stranded form, and these differences are detectable using gel electrophoresis in non-denaturing gels. In fSCCP, the oUgonucleotide primers are fluorescently labeled, which allows detection of the ampUmers in high- o throughput equipment such as DNA sequencing machines. Additionally, sequence database analysis methods, termed in siUco SNP (isSNP), are capable of identifying polymoφhisms by comparing the sequences of individual overlapping DNA fragments which assemble into a common consensus sequence. These computer-based methods filter out sequence variations due to laboratory preparation of DNA and sequencing errors using statistical models and automated analyses of DNA sequence chromatograms. In the alternative, SNPs may be detected and characterized by mass spectrometry using, for example, the high throughput MASSARRAY system (Sequenom, Inc., San Diego CA).
DNA-based identification techniques are critical in forensic technology. DNA sequences taken from very small biological samples such as tissues, e.g., hair or skin, or body fluids, e.g., blood, saliva, 5 semen, etc., can be ampUfied using, e.g., PCR, to identify individuals. (See, e.g., Erlich, H. (1992) PCR Technology, Freeman and Co., New York, NY). Similarly, polynucleotides of the present invention can be used as polymoφhic markers.
There is also a need for reagents capable of identifying the source of a particular tissue. Appropriate reagents can comprise, for example, DNA probes or primers prepared from the sequences 0 of the present invention that are specific for particular tissues. Panels of such reagents can identify tissue by species and/or by organ type. In a similar fashion, these reagents can be used to screen tissue cultures for contamination.
The polynucleotides of the present invention can also be used as molecular weight markers on nucleic acid gels or Southern blots, as diagnostic probes for the presence of a specific mRNA in a 5 particular cell type, in the creation of subtracted cDNA Ubraries which aid in the discovery of novel polynucleotides, in selection and synthesis of oUgomers for attachment to an array or other support, and as an antigen to elicit an immune response.
Disease Model Systems Using dithp o The dithp of the invention or their mammaUan homologs may be "knocked out" in an animal model system using homologous recombination in embryonic stem (ES) cells. Such techniques are well known in the art and are useful for the generation of animal models of human disease. (See, e.g., U.S. Patent Number 5,175,383 and U.S. Patent Number 5,767,337.) For example, mouse ES cells, such as the mouse 129/SvJ cell Une, are derived from the early mouse embryo and grown in culture. The ES 5 cells are transformed with a vector containing the gene of interest disrupted by a marker gene, e.g., the neomycin phosphotransferase gene (neo; Capecchi, M.R. (1989) Science 244:1288-1292). The vector integrates into the corresponding region of the host genome by homologous recombination. Aternatively, homologous recombination takes place using the Cre-loxP system to knockout a gene of interest in a tissue- or developmental stage-specific manner (Marth, J.D. (1996) Clin. Invest. 97:1999- o 2002; Wagner, K.U. et al. (1997) Nucleic Acids Res. 25 :4323-4330). Transformed ES cells are identified and microinjected into mouse cell blastocysts such as those from the C57BL/6 mouse strain. The blastocysts are surgically transferred to pseudopregnant dams, and the resulting chimeric progeny are genotyped and bred to produce heterozygous or homozygous strains. Transgenic animals thus generated may be tested with potential therapeutic or toxic agents. The dithp of the invention may also be manipulated in vitro in ES cells derived from human blastocysts. Human ES cells have the potential to differentiate into at least eight separate cell Uneages including endoderm, mesoderm, and ectodermal cell types. These cell Uneages differentiate into, for example, neural cells, hematopoietic Uneages, and cardiomyocytes (Thomson, J.A. et al. (1998) Science 5 282:1145-1147).
The dithp of the invention can also be used to create "knockin" humanized animals (pigs) or transgenic animals (mice or rats) to model human disease. With knockin technology, a region of dithp is injected into animal ES cells, and the injected sequence integrates into the animal cell genome. Transformed cells are injected into blastulae, and the blastulae are implanted as described above. o Transgenic progeny or inbred lines are studied and treated with potential pharmaceutical agents to obtain information on treatment of a human disease. Aternatively, a mammal inbred to overexpress dithp, resulting, e.g., in the secretion of DITHP in its milk, may also serve as a convenient source of that protein (Janne, J. et al. (1998) Biotechnol. Ainu. Rev. 4:55-74).
5 Screening Asavs
DITHP encoded by polynucleotides of the present invention may be used to screen for molecules that bind to or are bound by the encoded polypeptides. The binding of the polypeptide and the molecule may activate (agonist), increase, inhibit (antagonist), or decrease activity of the polypeptide or the bound molecule. Examples of such molecules include antibodies, oUgonucleotides, 0 proteins (e.g., receptors), or small molecules.
Preferably, the molecule is closely related to the natural Ugand of the polypeptide, e.g., a Ugand or fragment thereof, a natural substrate, or a structural or functional mimetic. (See, CoUgan et al., (1991) Current Protocols in Immunology 1(2): Chapter 5.) Similarly, the molecule can be closely related to the natural receptor to which the polypeptide binds, or to at least a fragment of the receptor, 5 e.g., the active site. In either case, the molecule can be rationally designed using known techniques.
Preferably, the screening for these molecules involves producing appropriate cells which express the polypeptide, either as a secreted protein or on the cell membrane. Preferred cells include cells from mammals, yeast, Drosophila, or E. coli. Cells expressing the polypeptide or cell membrane fractions which contain the expressed polypeptide are then contacted with a test compound and binding, o stimulation, or inhibition of activity of either the polypeptide or the molecule is analyzed.
Ai assay may simply test binding of a candidate compound to the polypeptide, wherein binding is detected by a fluorophore, radioisotope, enzyme conjugate, or other detectable label. Aternatively, the assay may assess binding in the presence of a labeled competitor. Additionally, the assay can be carried out using cell-free preparations, polypeptide/molecule affixed to a soUd support, chemical libraries, or natural product mixtures. The assay may also simply comprise the steps of mixing a candidate compound with a solution containing a polypeptide, measuring polypeptide/molecule activity or binding, and comparing the polypeptide/molecule activity or binding to 5 a standard.
Preferably, an ELISA assay using, e.g., a monoclonal or polyclonal antibody, can measure polypeptide level in a sample. The antibody can measure polypeptide level by either binding, directly or indirectly, to the polypeptide or by competing with the polypeptide for a substrate.
Al of the above assays can be used in a diagnostic or prognostic context. The molecules o discovered using these assays can be used to treat disease or to bring about a particular result in a patient (e.g., blood vessel growth) by activating or inhibiting the polypeptide/molecule. Moreover, the assays can discover agents which may inhibit or enhance the production of the polypeptide from suitably manipulated cells or tissues.
5 Transcript Imaging and Toxicological Testing
Aiother embodiment relates to the use of dithp to develop a transcript image of a tissue or cell type. A transcript image represents the global pattern of gene expression by a particular tissue or cell type. Global gene expression patterns are analyzed by quantifying the number of expressed genes and their relative abundance under given conditions and at a given time. (See Seilhamer et al., o "Comparative Gene Transcript Analysis," U.S. Patent Number 5,840,484, expressly incoφorated by reference herein.) Thus a transcript image may be generated by hybridizing the polynucleotides of the present invention or their complements to the totaUty of transcripts or reverse transcripts of a particular tissue or cell type. In one embodiment, the hybridization takes place in high-throughput format, wherein the polynucleotides of the present invention or their complements comprise a subset of a 5 plurality of elements on a microarray. The resultant transcript image would provide a profile of gene activity pertaining to human molecules for diagnostics and therapeutics.
Transcript images which profile dithp expression may be generated using transcripts isolated from tissues, cell Unes, biopsies, or other biological samples. The transcript image may thus reflect dithp expression in vivo, as in the case of a tissue or biopsy sample, or in vitro, as in the case of a cell 0 Une.
Transcript images which profile dithp expression may also be used in conjunction with in vitro model systems and preclinical evaluation of pharmaceuticals, as well as toxicological testing of industrial and naturally-occurring environmental compounds. Al compounds induce characteristic gene expression patterns, frequently termed molecular fingeφrints or toxicant signatures, which are indicative of mechanisms of action and toxicity (Nuwaysir, E. F. et al. (1999) Mol. Carcinog. 24:153- 159; Steiner, S. and Aiderson, N. L. (2000) Toxicol. Lett. 112-113:467-71, expressly incoφorated by reference herein). If a test compound has a signature similar to that of a compound with known toxicity, it is Ukely to share those toxic properties. These fingeφrints or signatures are most useful and 5 refined when they contain expression information from a large number of genes and gene famiUes. Ideally, a genome-wide measurement of expression provides the highest quality signature. Even genes whose expression is not altered by any tested compounds are important as well, as the levels of expression of these genes are used to normatize the rest of the expression data. The normatization procedure is useful for comparison of expression data after treatment with different compounds. While 0 the assignment of gene function to elements of a toxicant signature aids in inteφretation of toxicity mechanisms, knowledge of gene function is not necessary for the statistical matching of signatures which leads to prediction of toxicity. (See, for example, Press Release 00-02 from the National Institute of Environmental Health Sciences, released February 29, 2000, available at http://www.niehs.nih.gov/oc/news/toxchip.htm.) Therefore, it is important and desirable in 5 toxicological screening using toxicant signatures to include all expressed gene sequences.
In one embodiment, the toxicity of a test compound is assessed by treating a biological sample containing nucleic acids with the test compound. Nucleic acids that are expressed in the treated biological sample are hybridized with one or more probes specific to the polynucleotides of the present invention, so that transcript levels corresponding to the polynucleotides of the present o invention may be quantified. The transcript levels in the treated biological sample are compared with levels in an untreated biological sample. Differences in the transcript levels between the two samples are indicative of a toxic response caused by the test compound in the treated sample.
Aiother particular embodiment relates to the use of DITHP encoded by polynucleotides of the present invention to analyze the proteome of a tissue or cell type. The term proteome refers to the 5 global pattern of protein expression in a particular tissue or cell type. Each protein component of a proteome can be subjected individually to further analysis. Proteome expression patterns, or profiles, are analyzed by quantifying the number of expressed proteins and their relative abundance under given conditions and at a given time. A profile of a cell's proteome may thus be generated by separating and analyzing the polypeptides of a particular tissue or cell type. In one embodiment, the separation is o achieved using two-dimensional gel electtophoresis, in which proteins from a sample are separated by isoelectric focusing in the first dimension, and then according to molecular weight by sodium dodecyl sulfate slab gel electrophoresis in the second dimension (Steiner and Anderson, supra). The proteins are visuahzed in the gel as discrete and uniquely positioned spots, typically by staining the gel with an agent such as Coomassie Blue or silver or fluorescent stains. The optical density of each protein spot is generally proportional to the level of the protein in the sample. The optical densities of equivalently positioned protein spots from different samples, for example, from biological samples either treated or untreated with a test compound or therapeutic agent, are compared to identify any changes in protein spot density related to the treatment. The proteins in the spots are partially sequenced using, for example, standard methods employing chemical or enzymatic cleavage followed by mass spectiOmetry. The identity of the protein in a spot may be determined by comparing its partial sequence, preferably of at least 5 contiguous amino acid residues, to the polypeptide sequences of the present invention. In some cases, further sequence data may be obtained for definitive protein identification.
A proteomic profile may also be generated using antibodies specific for DITHP to quantify the levels of DITHP expression. In one embodiment, the antibodies are used as elements on a microarray, and protein expression levels are quantified by exposing the microarray to the sample and detecting the levels of protein bound to each array element (Lueking, A. et al. (1999) Anal. Biochem. 270:103-11; Mendoze, L. G. et al. (1999) Biotechniques 27:778-88). Detection may be performed by a variety of methods known in the art, for example, by reacting the proteins in the sample with a thiol- or amino- reactive fluorescent compound and detecting the amount of fluorescence bound at each array element.
Toxicant signatures at the proteome level are also useful for toxicological screening, and should be analyzed in parallel with toxicant signatures at the transcript level. There is a poor correlation between transcript and protein abundances for some proteins in some tissues (Anderson, N. L. and Seilhamer, J. (1997) Electrophoresis 18:533-537), so proteome toxicant signatures may be useful in the analysis of compounds which do not significantly affect the transcript image, but which alter the proteomic profile. In addition, the analysis of transcripts in body fluids is difficult, due to rapid degradation of mRNA, so proteomic profiling may be more reUable and informative in such cases. In another embodiment, the toxicity of a test compound is assessed by treating a biological sample containing proteins with the test compound. Proteins that are expressed in the treated biological sample are separated so that the amount of each protein can be quantified. The amount of each protein is compared to the amount of the corresponding protein in an untreated biological sample. A difference in the amount of protein between the two samples is indicative of a toxic response to the test compound in the treated sample. Individual proteins are identified by sequencing the amino acid residues of the individual proteins and comparing these partial sequences to the DITHP encoded by polynucleotides of the present invention.
In another embodiment, the toxicity of a test compound is assessed by treating a biological sample containing proteins with the test compound. Proteins from the biological sample are incubated with antibodies specific to the DITHP encoded by polynucleotides of the present invention. The amount of protein recognized by the antibodies is quantified. The amount of protein in the treated biological sample is compared with the amount in an untreated biological sample. A difference in the amount of protein between the two samples is indicative of a toxic response to the test compound in the treated sample.
Transcript images may be used to profile dithp expression in distinct tissue types. This process 5 can be used to determine human molecule activity in a particular tissue type relative to this activity in a different tissue type. Transcript images may be used to generate a profile of dithp expression characteristic of diseased tissue. Transcript images of tissues before and after treatment may be used for diagnostic puφoses, to monitor the progression of disease, and to monitor the efficacy of drug treatments for diseases which affect the activity of human molecules. o Transcript images of cell Unes can be used to assess human molecule activity and/or to identify cell lines that lack or misregulate this activity. Such cell Unes may then be treated with pharmaceutical agents, and a transcript image following treatment may indicate the efficacy of these agents in restoring desired levels of this activity. A similar approach may be used to assess the toxicity of pharmaceutical agents as reflected by undesirable changes in human molecule activity. Candidate pharmaceutical 5 agents may be evaluated by comparing their associated transcript images with those of pharmaceutical agents of known effectiveness.
Antisense Molecules
The polynucleotides of the present invention are useful in antisense technology. Antisense o technology or therapy reUes on the modulation of expression of a target protein through the specific binding of an antisense sequence to a target sequence encoding the target protein or directing its expression. (See, e.g., Agrawal, S., ed. (1996) Antisense Therapeutics, Humana Press Inc., Totawa NJ; Aa a, A. et al. (1997) Pharmacol. Res. 36(3):171-178; Crooke, S.T. (1997) Adv. Pharmacol. 40:1-49; Sharma, H.W. and R. Narayanan (1995) Bioessays 17(12):1055-1063; andLavrosky, Y. et 5 al. (1997) Biochem Mol. Med. 62(1):11-22.) An antisense sequence is a polynucleotide sequence capable of specifically hybridizing to at least a portion of the target sequence. Antisense sequences bind to cellular mRNA and/or genomic DNA, affecting translation and/or transcription. Antisense sequences can be DNA, RNA, or nucleic acid mimics and analogs. (See, e.g., Rossi, J.J. et al. (1991) Antisense Res. Dev. l(3):285-288; Lee, R. et al. (1998) Biochemistry 37(3):900-1010; Pardridge, 0 W.M. et al. (1995) Proc. Natl. Acad. Sci. USA 92(12):5592-5596; and Nielsen, P. E. and Haaima, G. (1997) Chem. Soc. Rev. 96:73-78.) Typically, the binding which results in modulation of expression occurs through hybridization or binding of complementary base pairs. Antisense sequences can also bind to DNA duplexes through specific interactions in the major groove of the double helix. The polynucleotides of the present invention and fragments thereof can be used as antisense sequences to modify the expression of the polypeptide encoded by dithp. The antisense sequences can be produced ex vivo, such as by using any of the ABI nucleic acid synthesizer series (PE Biosystems) or other automated systems known in the art. Antisense sequences can also be produced biologically, such as by transforming an appropriate host cell with an expression vector containing the sequence of interest. (See, e.g., Agrawal, supra.)
In therapeutic use, any gene delivery system suitable for introduction of the antisense sequences into appropriate target cells can be used. Antisense sequences can be delivered inttacellularly in the form of an expression plasmid which, upon transcription, produces a sequence complementary to at least a portion of the cellular sequence encoding the target protein. (See, e.g., Slater, J.E., et al. (1998) J. Alergy CUa Immunol. 102(3):469-475; and Scanlon, K.J., et al. (1995) 9(13):1288-1296.) Aitisense sequences can also be introduced inttacellularly through the use of viral vectors, such as retrovirus and adeno-associated virus vectors. (See, e.g., Miller, A.D. (1990) Blood 76:271; Ausubel, F.M. et al. (1995) Current Protocols in Molecular Biology. John Wiley & Sons, New York NY; Uckert, W. and W. Walther (1994) Pharmacol. Ther. 63(3):323-347.) Other gene delivery mechanisms include liposome-derived systems, artificial viral envelopes, and other systems known in the art. (See, e.g., Rossi, J.J. (1995) Br. Med. Bull. 51(l):217-225; Boado, R.J. et al. (1998) J. Pharra Sci. 87(11):1308- 1315; and Morris, M.C. et al. (1997) Nucleic Acids Res. 25(14):2730-2736.)
Expression
In order to express a biologically active DITHP, the nucleotide sequences encoding DITHP or fragments thereof may be inserted into an appropriate expression vector, i.e., a vector which contains the necessary elements for transcriptional and translational control of the inserted coding sequence in a suitable host. Methods which are well known to those skilled in the art may be used to construct expression vectors containing sequences encoding DITHP and appropriate transcriptional and translational control elements. These methods include in vitro recombinant DNA techniques, synthetic techniques, and in vivo genetic recombination. (See, e.g., Sambrook, supra. Chapters 4, 8, 16, and 17; and Ausubel, supra, Chapters 9, 10, 13, and 16.)
A variety of expression vector/host systems may be utiUzed to contain and express sequences encoding DITHP. These include, but are not Umited to, microorganisms such as bacteria ttansformed with recombinant bacteriophage, plasmid, or cosmid DNA expression vectors; yeast ttansformed with yeast expression vectors; insect cell systems infected with viral expression vectors (e.g., baculovirus); plant cell systems transformed with viral expression vectors (e.g., cauliflower mosaic virus, CaMV, or tobacco mosaic virus, TMV) or with bacterial expression vectors (e.g., Ti or pBR322 plasmids); or animal (mammalian) cell systems. (See, e.g., Sambrook, supra; Ausubel, 1995, supra. Van Heeke, G. and S.M. Schuster (1989) J. Biol. Chem. 264:5503-5509; Bitter, G.A et al. (1987) Methods Enzymol. 153:516-544; Scorer, CA. et al. (1994) Bio/Technology 12:181-184; Engelhard, E.K. et al. (1994) Proc. Natl. Acad. Sci. USA 91:3224-3227; Sandig, V. et al. (1996) Hum. Gene Ther. 7:1937-1945; 5 Takamatsu, N. (1987) EMBO J. 6:307-311 ; Coruzzi, G. et al. (1984) EMBO J. 3:1671-1680; BrogUe, R. et al. (1984) Science 224:838-843; Winter, J. et al. (1991) Results Probl. Cell Differ. 17:85-105; The McGraw Hill Yearbook of Science and Technology (1992) McGraw Hill, New York NY, pp. 191-196; Logan, J. and T. Shenk (1984) Proc. Natl. Acad. Sci. USA 81:3655-3659; and Harrington, J.J. et al. (1997) Nat. Genet. 15:345-355.) Expression vectors derived from rettoviruses, adenoviruses, 0 or heφes or vaccinia viruses, or from various bacterial plasmids, may be used for deUvery of nucleotide sequences to the targeted organ, tissue, or cell population. (See, e.g., Di Nicola, M. et al. (1998) Cancer Gen. Ther. 5(6):350-356; Yu, M. et al., (1993) Proc. Natl. Acad. Sci. USA 90(13):6340-6344; Buller, R.M. et al. (1985) Nature 317(6040):813-815; McGregor, D.P. et al. (1994) Mol. Immunol. 31(3):219-226; and Verma, I.M. and N. Somia (1997) Nature 389:239-242.) The invention is not 5 Umited by the host cell employed.
For long term production of recombinant proteins in mammaUan systems, stable expression of DITHP in cell Unes is preferred. For example, sequences encoding DITHP can be ttansformed into cell Unes using expression vectors which may contain viral origins of repUcation and/or endogenous expression elements and a selectable marker gene on the same or on a separate vector. Any number of o selection systems may be used to recover transformed cell Unes. (See, e.g., Wigler, M. et al. (1977)
Cell 11:223-232; Lowy, I. et al. (1980) Cell 22:817-823.; Wigler, M. et al. (1980) Proc. Natl. Acad. Sci. USA 77:3567-3570; Colbere-Garapin, F. et al. (1981) J. Mol. Biol. 150:1-14; Hartman, S.C and R.CMulligan (1988) Proc. Nail. Acad. Sci. USA 85:8047-8051; Rhodes, CA (1995) Methods Mol. Biol. 55:121-131.) 5
Therapeutic Uses of dithp
The dithp of the invention may be used for somatic or germUne gene therapy. Gene therapy may be performed to (i) correct a genetic deficiency (e.g., in the cases of severe combined immunodeficiency (SCID)-Xl disease characterized by X-Unked inheritance (Cavazzana-Calvo, M. et o al. (2000) Science 288 :669-672), severe combined immunodeficiency syndrome associated with an inherited adenosine deaminase (ADA) deficiency (Blaese, R.M. et al. (1995) Science 270:475-480; Bordignon, C et al. (1995) Science 270:470-475), cystic fibrosis (Zabner, J. et al. (1993) Cell 75:207- 216; Crystal, R.G. et al. (1995) Hum. Gene Therapy 6:643-666; Crystal, R.G. et al. (1995) Hum. Gene Therapy 6:667-703), thalassemias, famiUal hypercholesterolemia, and hemophiUa resulting from Factor VIII or Factor LX deficiencies (Crystal, R.G. (1995) Science 270:404-410; Verma, I.M. and Somia, N. (1997) Nature 389:239-242)), (u) express a conditionally lethal gene product (e.g., in the case of cancers which result from unregulated cell proUferation), or (Ui) express a protein which affords protection against intracellular parasites (e.g., against human retroviruses, such as human 5 immunodeficiency virus (HIV) (Baltimore, D. (1988) Nature 335:395-396; Poeschla, E. et al. (1996) Proc. Natl. Acad. Sci. USA. 93:11395-11399), hepatitis B or C virus (HBV, HCV); fungal parasites, such as Candida albicans and Paracoccidioides brasihensis; and protozoan parasites such as Plasmodium falciparum and Trvpanosoma cruzi). In the case where a genetic deficiency in dithp expression or regulation causes disease, the expression of dithp from an appropriate population of o transduced cells may alleviate the clinical manifestations caused by the genetic deficiency.
In a further embodiment of the invention, diseases or disorders caused by deficiencies in dithp are treated by constructing mammaUan expression vectors comprising dithp and introducing these vectors by mechanical means into dithp-deficient cells. Mechanical transfer technologies for use with cells in vivo or ex vitro include (i) direct DNA microinjection into individual cells, (u) balUstic gold 5 particle delivery, (ui) Uposome-mediated transfection, (iv) receptor-mediated gene transfer, and (v) the use of DNA transposons (Morgan, R.A. and Anderson, W.F. (1993) Annu. Rev. Biochem. 62:191-217; Ivies, Z. (1997) Cell 91:501-510; Boulay, J-L. and Rέcipon, H. (1998) Curr. Opin. Biotechnol. 9:445- 450).
Expression vectors that may be effective for the expression of dithp include, but are not Umited o to, the PCDNA 3.1, EPITAG, PRCCMV2, PREP, PVAX vectors (Invitrogen, Carlsbad C A),
PCMV-SCRIPT, PCMV-TAG, PEGSH/PERV (Stratagene, La Jolla CA), and PTET-OFF, PTET-ON, PTRE2, PTRE2-LUC, PTK-HYG (Clontech, Palo Ato CA). The dithp of the invention may be expressed using (i) a constitutively active promoter, (e.g., from cytomegalovirus (CMV), Rous sarcoma virus (RSV), SV40 virus, thymidine kinase (TK), or β-actin genes), (n) an inducible promoter 5 (e.g., the tetracycUne-regulated promoter (Gossen, M. and Bujard, H. (1992) Proc. Natl. Acad. Sci. U.S.A. 89:5547-5551; Gossen, M. et al., (1995) Science 268:1766-1769; Rossi, F.M.V. and Blau, H.M. (1998) Curr. Opin. Biotechnol. 9:451-456), commercially available in the T-REX plasmid (Invitrogen); the ecdysone-inducible promoter (available in the plasmids PVGRXR and PIND; Invitrogen); the FK506/rapamycin inducible promoter; or the RU486/mifepristone inducible promoter 0 (Rossi, F.M.V. and Blau, H.M. supra), or (Ui) a tissue-specific promoter or the native promoter of the endogenous gene encoding DITHP from a normal individual.
Commercially available Uposome transformation kits (e.g., the PERFECT LIPID TRANSFECTION KIT, available from Invitrogen) allow one with ordinary skill in the art to deliver polynucleotides to target cells in culture and require minimal effort to optimize experimental parameters. In the alternative, transformation is performed using the calcium phosphate method (Graham, F.L. andEb, AJ. (1973) Virology 52:456-467), or by electroporation (Neumann, E. et al. (1982) EMBO J. 1:841-845). The introduction of DNA to primary cells requires modification of these standardized mammaUan transfection protocols. 5 In another embodiment of the invention, diseases or disorders caused by genetic defects with respect to dithp expression are treated by constructing a rettovirus vector consisting of (i) dithp under the control of an independent promoter or the rettovirus long terminal repeat (LTR) promoter, (U) appropriate RNA packaging signals, and (iti) a Rev-responsive element (RRE) along with additional rettovirus cis-acting RNA sequences and coding sequences required for efficient vector propagation. 0 Rettovirus vectors (e.g., PFB and PFBNEO) are commercially available (Sttatagene) and are based on pubUshed data (Riviere, I. et al. (1995) Proc. Natl. Acad. Sci. U.S.A. 92:6733-6737), incoφorated by reference herein. The vector is propagated in an appropriate vector producing cell Une (VPCL) that expresses an envelope gene with a tropism for receptors on the target cells or a promiscuous envelope protein such as VSVg (Amentano, D. et al. (1987) J. Virol. 61:1647-1650; Bender, M.A. et al. (1987) 5 J. Virol. 61 :1639-1646; Adam, M.A and Miller, AD. (1988) J. Virol. 62:3802-3806; Dull, T. et al. (1998) J. Virol. 72:8463-8471; Zufferey, R. et al. (1998) J. Virol. 72:9873-9880). U.S. Patent Number 5,910,434 to Rigg ("Method for obtaining retrovirus packaging cell Unes producing high transducing efficiency retroviral supernatant") discloses a method for obtaining retrovirus packaging cell Unes and is hereby incoφorated by reference. Propagation of retrovirus vectors, transduction of a population of o cells (e.g., CD4+ T-cells), and the return of ttansduced cells to a patient are procedures well known to persons skilled in the art of gene therapy and have been well documented (Ranga, U. et al. (1997) J. Virol. 71:7020-7029; Bauer, G. et al. (1997) Blood 89:2259-2267; Bonyhadi, M.L. (1997) J. Virol. 71:4707-4716; Ranga, U. et al. (1998) Proc. Natl. Acad. Sci. U.S.A. 95:1201-1206; Su, L. (1997) Blood 89:2283-2290). 5 In the alternative, an adenovirus-based gene therapy delivery system is used to deUver dithp to cells which have one or more genetic abnormaUties with respect to the expression of dithp. The construction and packaging of adenovirus-based vectors are well known to those with ordinary skill in the art. RepUcation defective adenovirus vectors have proven to be versatile for importing genes encoding immunoregulatory proteins into intact islets in the pancreas (Csete, M.E. et al. (1995) o Transplantation 27:263-268). Potentially useful adenoviral vectors are described in U.S. Patent
Number 5,707,618 to Amentano ("Adenovirus vectors for gene therapy"), hereby incoφorated by reference. For adenoviral vectors, see also Antinozzi, P.A et al. (1999) Ainu. Rev. Nutr. 19:511-544 and Verma, I.M. and Somia, N. (1997) Nature 18:389:239-242, both incoφorated by reference herein. In another alternative, a herpes-based, gene therapy deUvery system is used to deliver dithp to target cells which have one or more genetic abnormatities with respect to the expression of dithp. The use of heφes simplex virus (HSV)-based vectors may be especially valuable for introducing dithp to cells of the central nervous system, for which HSV has a tropism. The construction and packaging of 5 heφes-based vectors are well known to those with ordinary skill in the art. A repUcation-competent heφes simplex virus (HSV) type 1 -based vector has been used to deUver a reporter gene to the eyes of primates (Liu, X. et al. (1999) Exp. Eye Res.l69:385-395). The construction of a HSV-1 virus vector has also been disclosed in detail in U.S. Patent Number 5,804,413 to DeLuca ("Heφes simplex virus strains for gene transfer"), which is hereby incoφorated by reference. U.S. Patent Number 5,804,413 o teaches the use of recombinant HSV d92 which consists of a genome containing at least one exogenous gene to be transferred to a cell under the control of the appropriate promoter for puφoses including human gene therapy. Aso taught by this patent are the construction and use of recombinant HSV strains deleted for ICP4, ICP27 and ICP22. For HSV vectors, see also Goins, W. F. et al. 1999 J. Virol. 73:519-532 and Xu, H. et al., (1994) Dev. Biol. 163:152-161, hereby incoφorated by reference. 5 The manipulation of cloned heφesvirus sequences, the generation of recombinant virus following the transfection of multiple plasmids containing different segments of the large heφesvirus genomes, the growth and propagation of heφesvirus, and the infection of cells with heφesvirus are techniques well known to those of ordinary skill in the art.
In another alternative, an alphavirus (positive, single-stranded RNA virus) vector is used to o deUver dithp to target cells. The biology of the prototypic alphavirus, SemUki Forest Virus (SFV), has been studied extensively and gene transfer vectors have been based on the SFV genome (Garoff, H. and Li, K-J. (1998) Curr. Opin. Biotech. 9:464-469). During alphavirus RNA repUcation, a subgenomic RNA is generated that normally encodes the viral capsid proteins. This subgenomic RNA repUcates to higher levels than the full-length genomic RNA, resulting in the overproduction of capsid proteins 5 relative to the viral proteins with enzymatic activity (e.g., protease and polymerase). Similarly, inserting dithp into the alphavirus genome in place of the capsid-coding region results in the production of a large number of dithp RNA and the synthesis of high levels of DITHP in vector transduced cells. While alphavirus infection is typically associated with cell lysis within a few days, the abiUty to estabUsh a persistent infection in hamster normal kidney cells (BHK-21) with a variant of Sindbis virus o (SIN) indicates that the lytic repUcation of alphaviruses can be altered to suit the needs of the gene therapy application (Dryga, S.A. et al. (1997) Virology 228:74-83). The wide host range of alphaviruses will allow the introduction of dithp into a variety of cell types. The specific transduction of a subset of cells in a population may require the sorting of cells prior to transduction. The methods of manipulating infectious cDNA clones of alphaviruses, performing alphavirus cDNA and RNA ttansfections, and performing alphavirus infections, are well known to those with ordinary skill in the art.
Antibodies 5 Anti-DITHP antibodies may be used to analyze protein expression levels. Such antibodies include, but are not limited to, polyclonal, monoclonal, chimeric, single chain, and Fab fragments. For descriptions of and protocols of antibody technologies, see, e.g., Pound J.D. (1998) Immunochemical Protocols, Humana Press, Totowa, NJ.
The amino acid sequence encoded by the dithp of the Sequence Listing may be analyzed by o appropriate software (e.g. , LASERGENE NAVIGATOR software, DNASTAR) to determine regions of high immunogenicity. The optimal sequences for immunization are selected from the C-terminus, the N-terminus, and those intervening, hydrophiUc regions of the polypeptide which are likely to be exposed to the external environment when the polypeptide is in its natural conformation. Analysis used to select appropriate epitopes is also described by Ausubel (1997, supra. Chapter 11.7). Peptides used for 5 antibody induction do not need to have biological activity; however, they must be antigenic. Peptides used to induce specific antibodies may have an amino acid sequence consisting of at five amino acids, preferably at least 10 amino acids, and most preferably 15 amino acids. A peptide which mimics an antigenic fragment of the natural polypeptide may be fused with another protein such as keyhole Umpet cyanin (KLH; Sigma, St. Louis MO) for antibody production. A peptide encompassing an antigenic o region may be expressed from a dithp, synthesized as described above, or purified from human cells.
Procedures well known in the art may be used for the production of antibodies. Various hosts including mice, goats, and rabbits, may be immunized by injection with a peptide. Depending on the host species, various adjuvants may be used to increase immunological response.
In one procedure, peptides about 15 residues in length may be synthesized using an ABI 431 A 5 peptide synthesizer (PE Biosystems) using fmoc-chemistry and coupled to KLH (Sigma) by reaction with M-maleimidobenzoyl-N-hydroxysuccinimide ester (Ausubel, 1995, supra). Rabbits are immunized with the peptide-KLH complex in complete Freund's adjuvant. The resulting antisera are tested for antipeptide activity by binding the peptide to plastic, blocking with 1 % bovine serum albumin (BSA), reacting with rabbit antisera, washing, and reacting with radioiodinated goat anti-rabbit IgG. o Antisera with antipeptide activity are tested for anti-DITHP activity using protocols well known in the art, including ELISA, radioimmunoassay (RIA), and immunoblotting.
In another procedure, isolated and purified peptide may be used to immunize mice (about 100 μg of peptide) or rabbits (about 1 mg of peptide). Subsequently, the peptide is radioiodinated and used to screen the immunized animals' B-lymphocytes for production of antipeptide antibodies. Positive cells are then used to produce hybridomas using standard techniques. About 20 mg of peptide is sufficient for labeUng and screening several thousand clones. Hybridomas of interest are detected by screening with radioiodinated peptide to identify those fusions producing peptide-specific monoclonal antibody. In a typical protocol, wells of a multi-well plate (FAST, Becton-Dickinson, Palo Ato, CA) 5 are coated with affinity-purified, specific rabbit-anti-mouse (or suitable anti-species IgG) antibodies at 10 mg/ml. The coated wells are blocked with 1 % BSA and washed and exposed to supematants from hybridomas. After incubation, the wells are exposed to radiolabeled peptide at 1 mg/ml.
Clones producing antibodies bind a quantity of labeled peptide that is detectable above background. Such clones are expanded and subjected to 2 cycles of cloning. Cloned hybridomas are o injected into pristane-tteated mice to produce ascites, and monoclonal antibody is purified from the ascitic fluid by affinity chromatography on protein A (Amersham Pharmacia Biotech). Several procedures for the production of monoclonal antibodies, including in vitro production, are described in Pound (supra). Monoclonal antibodies with antipeptide activity are tested for anti-DITHP activity using protocols well known in the art, including ELISA, RIA, and immunoblotting. 5 Aitibody fragments containing specific binding sites for an epitope may also be generated. For example, such fragments include, but are not Umited to, the F(ab')2 fragments produced by pepsin digestion of the antibody molecule, and the Fab fragments generated by reducing the disulfide bridges of the F(ab')2 fragments. Aternatively, construction of Fab expression libraries in filamentous bacteriophage allows rapid and easy identification of monoclonal fragments with desired specificity 0 (Pound, supra. Chaps. 45-47). Aitibodies generated against polypeptide encoded by dithp can be used to purify and characterize full-length DITHP protein and its activity, binding partners, etc.
Asavs Using Aitibodies
Aiti-DITHP antibodies may be used in assays to quantify the amount of DITHP found in a 5 particular human cell. Such assays include methods utiUzing the antibody and a label to detect expression level under normal or disease conditions. The peptides and antibodies of the invention may be used with or without modification or labeled by joining them, either covalently or noncovalently, with a reporter molecule.
Protocols for detecting and measuring protein expression using either polyclonal or monoclonal o antibodies are well known in the art. Examples include ELISA, RIA, and fluorescent activated cell sorting (FACS). Such immunoassays typically involve the formation of complexes between the DITHP and its specific antibody and the measurement of such complexes. These and other assays are described in Pound (supra). Without further elaboration, it is beUeved that one skilled in the art can, using the preceding description, utiUze the present invention to its fullest extent. The following preferred specific embodiments are, therefore, to be construed as merely illusttative, and not Umitative of the remainder of the disclosure in any way whatsoever. The disclosures of all patents, appUcations, and pubUcations mentioned above and below, in particular U.S. Ser. No. 60/156,294, U.S. Ser. No. 60/155,760, U.S. Ser. No. 60/155,939, U.S. Ser. No. 60/156,565, U.S. Ser. No. 60/156,624, U.S. Ser. No. 60/156,625, U.S. Ser. No. 60/167,542, U.S. Ser. No. 60/167,522, U.S. Ser. No. 60/167,453, U.S. Ser. No. 60/167,517, U.S. Ser. No. 60/167,943, U.S. Ser. No. 60/167,945, U.S. Ser. No. 60/167,520, U.S. Ser. No. 60/168,468, U.S. Ser. No. 60/168,599, U.S. Ser. No. 60/167,410, U.S. Ser. No. 60/168,265, U.S. Ser. No. 60/168,429, U.S. Ser. No. 60/168,432, U.S. Ser. No. 60/167,521, U.S. Ser. No. 60/168,857, U.S. Ser. No. 60/168,197, U.S. Ser. No. 60/168,611, and U.S. Ser. No. 60/168,613 are hereby expressly incoφorated by reference.
EXAMPLES I. Construction of cDNA Libraries
RNA was purchased from CLONTECH Laboratories, Inc. (Palo Ato CA) or isolated from various tissues. Some tissues were homogenized and lysed in guanidinium isothiocyanate, while others were homogenized and lysed in phenol or in a suitable mixture of denaturants, such as TRIZOL (Life Technologies), a monophasic solution of phenol and guanidine isothiocyanate. The resulting lysates were centrifuged over CsCl cushions or extracted with chloroform. RNA was precipitated with either isopropanol or sodium acetate and ethanol, or by other routine methods.
Phenol extraction and precipitation of RNA were repeated as necessary to increase RNA purity. In most cases, RNA was treated with DNase. For most Ubraries, poly(A+) RNA was isolated using oUgo d(T)-coupled paramagnetic particles (Promega Coφoration (Promega), Madison WI), OLIGOTEX latex particles (QIAGEN, Inc. (QIAGEN), Valencia CA), or an OLIGOTEX mRNA purification kit (QIAGEN). Aternatively, RNA was isolated directly from tissue lysates using other RNA isolation kits, e.g., the POLY(A)PURE mRNA purification kit (Ambion, Inc., Austin TX).
In some cases, Stratagene was provided with RNA and constructed the corresponding cDNA Ubraries. Otherwise, cDNA was synthesized and cDNA Ubraries were constructed with the UNLZAP vector system (Stratagene Cloning Systems, Inc. (Stratagene), La Jolla CA) or SUPERSCRIPT plasmid system (Life Technologies), using the recommended procedures or similar methods known in the art. (See, e.g., Ausubel, 1997, supra. Chapters 5.1 through 6.6.) Reverse transcription was initiated using oligo d(T) or random primers. Synthetic oligonucleotide adapters were Ugated to double stranded cDNA, and the cDNA was digested with the appropriate restriction enzyme or enzymes. For most Ubraries, the cDNA was size-selected (300-1000 bp) using SEPHACRYL S1000, SEPHAROSE CL2B, or SEPHAROSE CL4B column chromatography (Amersham Pharmacia Biotech) or preparative agarose gel electtophoresis. cDN A were Ugated into compatible restriction enzyme sites of the polylinker of a suitable plasmid, e.g., PBLUESCRIPT plasmid (Stratagene), pSPORTl plasmid 5 (Life Technologies), or pINCY (Incyte). Recombinant plasmids were transformed into competent E. coU cells including XLl-Blue, XLl-BlueMRF, or SOLR from Stratagene or DH5α, DH10B, or ElectroMAX DH10B from Life Technologies.
II. Isolation of cDNA Clones o Plasmids were recovered from host cells by in vivo excision using the UNIZAP vector system
(Stratagene) or by cell lysis. Plasmids were purified using at least one of the following: the Magic or WIZARD Minipreps DNA purification system (Promega); the AGTC Miniprep purification kit (Edge BioSystems, Gaithersburg MD); and the QIAWELL 8, QIAWELL 8 Plus, and QIAWELL 8 Ultra plasmid purification systems or the R.E. A.L. PREP 96 plasmid purification kit (QIAGEN). Following 5 precipitation, plasmids were resuspended in 0.1 ml of distilled water and stored, with or without lyophiUzation, at 4°C
Aternatively, plasmid DNA was amplified from host cell lysates using direct Unk PCR in a high-throughput format. (Rao, V.B. (1994) Anal. Biochem. 216:1-14.) Host cell lysis and thermal cycling steps were carried out in a single reaction mixture. Samples were processed and stored in 384- o well plates, and the concentration of ampUfied plasmid DNA was quantified fluorometticaUy using
PICOGREEN dye (Molecular Probes, Inc. (Molecular Probes), Eugene OR) and a FLUOROSKAN II fluorescence scanner (Labsystems Oy, Helsinki, Finland).
III. Sequencing and Analysis 5 cDNA sequencing reactions were processed using standard methods or high-throughput instrumentation such as the ABI CATALYST 800 thermal cycler (PE Biosystems) or the PTC-200 thermal cycler (MJ Research) in conjunction with the HYDRA microdispenser (Robbins Scientific Coφ., Sunnyvale CA) or the MICROLAB 2200 Uquid transfer system (Hamilton). cDNA sequencing reactions were prepared using reagents provided by Amersham Pharmacia Biotech or suppUed in ABI o sequencing kits such as the ABI PRISM BIGDYE Terminator cycle sequencing ready reaction kit (PE
Biosystems). Electtophoretic separation of cDNA sequencing reactions and detection of labeled polynucleotides were carried out using the MEGABACE 1000 DNA sequencing system (Molecular Dynamics); the ABI PRISM 373 or 377 sequencing system (PE Biosystems) in conjunction with standard ABI protocols and base calting software; or other sequence analysis systems known in the art. Reading frames within the cDNA sequences were identified using standard methods (reviewed in Ausubel, 1997, supra. Chapter 7.7). Some of the cDNA sequences were selected for extension using the techniques disclosed in Example VIII.
IV. Assembly and Analysis of Sequences
Component sequences from chromatograms were subject to PHRED analysis and assigned a quaUty score. The sequences having at least a required quatity score were subject to various preprocessing editing pathways to eUminate, e.g., low quaUty 3' ends, vector and linker sequences, polyA tails, Au repeats, mitochondrial and ribosomal sequences, bacterial contamination sequences, and sequences smaller than 50 base pairs. In particular, low-information sequences and repetitive elements (e.g., dinucleotide repeats, Au repeats, etc.) were replaced by "n's", or masked, to prevent spurious matches.
Processed sequences were then subject to assembly procedures in which the sequences were assigned to gene bins (bins). Each sequence could only belong to one bin. Sequences in each gene bin were assembled to produce consensus sequences (templates). Subsequent new sequences were added to existing bins using BLASTn (v.1.4 WashU) and CROSSMATCH. Candidate pairs were identified as all BLAST hits having a quaUty score greater than or equal to 150. Aignments of at least 82% local identity were accepted into the bin. The component sequences from each bin were assembled using a version of PHRAP. Bins with several overlapping component sequences were assembled using DEEP PHRAP. The orientation (sense or antisense) of each assembled template was determined based on the number and orientation of its component sequences. Template sequences as disclosed in the sequence Usting correspond to sense strand sequences (the "forward" reading frames), to the best determination. The complementary (antisense) strands are inherently disclosed herein. The component sequences which were used to assemble each template consensus sequence are listed in Table 4, along with their positions along the template nucleotide sequences.
Bins were compared against each other and those having local similarity of at least 82% were combined and reassembled. Reassembled bins having templates of insufficient overlap (less than 95% local identity) were re-split. Asembled templates were also subject to analysis by STITCHER/EXON MAPPER algorithms which analyze the probabilities of the presence of spUce variants, alternatively spUced exons, spUce junctions, differential expression of alternative spliced genes across tissue types or disease states, etc. These resulting bins were subject to several rounds of the above assembly procedures.
Once gene bins were generated based upon sequence aUgnments, bins were clone joined based upon clone informatioa If the 5' sequence of one clone was present in one bin and the 3' sequence from the same clone was present in a different bin, it was Ukely that the two bins actually belonged together in a single bin. The resulting combined bins underwent assembly procedures to regenerate the consensus sequences.
The final assembled templates were subsequently annotated using the following procedure. 5 Template sequences were analyzed using BLASTn (v2.0, NCBI) versus gbpri (GenBank version 118). "Hits" were defined as an exact match having from 95% local identity over 200 base pairs through 100% local identity over 100 base pairs, or a homolog match having an E-value, i.e. a probability score, of < 1 x IO"8. The hits were subject to frameshift FASTx versus GENPEPT (GenBank version 118). (See Table 6). In this analysis, a homolog match was defined as having an E-value of ≤ 1 x IO"8. 0 The assembly method used above was described in "System and Methods for Analyzing Biomolecular Sequences," U.S.S.N. 09/276,534, filed March 25, 1999, and the LIFESEQ Gold user manual (Incyte) both incoφorated by reference herein.
Following assembly, template sequences were subjected to motif, BLAST, and functional analyses, and categorized in protein hierarchies using methods described in, e.g., "Database System 5 Employing Protein Function Hierarchies for Viewing Biomolecular Sequence Data," U.S.S.N. 08/812,290, filed March 6, 1997; "Relational Database for Storing Biomolecule Information," U.S.S.N. 08/947,845, filed October 9, 1997; "Project-Based Full-Length Biomolecular Sequence Database," U.S.S.N. 08/811,758, filed March 6, 1997; and "Relational Database and System for Storing Information Relating to Biomolecular Sequences," U.S.S.N. 09/034,807, filed March 4, 1998, 0 all of which are incoφorated by reference herein.
The template sequences were further analyzed by translating each template in all three forward reading frames and searching each translation against the Pfam database of hidden Markov model- based protein famiUes and domains using the HMMER software package (available to the pubUc from Washington University School of Medicine, St. Louis MO). Regions of templates which, when 5 translated, contain similarity to Pfam consensus sequences are reported in Table 2, along with descriptions of Pfam protein domains and families. Only those Pfam hits with an E-value of ≤ 1 x IO"3 are reported. (See also World Wide Web site http://pfarawustl.edu/ for detailed descriptions of Pfam protein domains and famiUes.)
Additionally, the template sequences were translated in all three forward reading frames, and o each translation was searched against hidden Markov models for signal peptide and ttansmembrane domains using the HMMER software package. Construction of hidden Markov models and their usage in sequence analysis has been described. (See, for example, Eddy, S.R. (1996) Curr. Opin. Str. Biol. 6:361-365.) Regions of templates which, when translated, contain similarity to signal peptide or transmembrane domain consensus sequences are reported in Table 3. Only those signal peptide or ttansmembrane hits with a cutoff score of 11 bits or greater are reported. A cutoff score of 11 bits or greater corresponds to at least about 91-94% true-positives in signal peptide prediction, and at least about 75% true-positives in transmembrane domain prediction.
The results of HMMER analysis as reported in Tables 2 and 3 may support the results of BLAST analysis as reported in Table 1 or may suggest alternative or additional properties of template- encoded polypeptides not previously uncovered by BLAST or other analyses.
Template sequences are further analyzed using the bioinformatics tools Usted in Table 6, or using sequence analysis software known in the art such as MACDNASIS PRO software (Hitachi Software Engineering, South San Francisco CA) and LASERGENE software (DNASTAR). Template sequences may be further queried against pubUc databases such as the GenBank rodent, mammaUan, vertebrate, prokaryote, and eukaryote databases.
V. Analysis of Polynucleotide Expression
Northern analysis is a laboratory technique used to detect the presence of a transcript of a gene and involves the hybridization of a labeled nucleotide sequence to a membrane on which RN from a particular cell type or tissue have been bound. (See, e.g., Sambrook, supra, ch. 7; Ausubel, 1995, supra, ch. 4 and 16.)
Analogous computer techniques applying BLAST were used to search for identical or related molecules in cDNA databases such as GenBank or LIFESEQ (Incyte Genomics). This analysis is much faster than multiple membrane-based hybridizations. In addition, the sensitivity of the computer search can be modified to determine whether any particular match is categorized as exact or similar. The basis of the search is the product score, which is defined as:
BLAST Score x Percent Identity 5 x minimum {length(Seq. 1), length(Seq. 2)}
The product score takes into account both the degree of similarity between two sequences and the length of the sequence match. The product score is a normaUzed value between 0 and 100, and is calculated as follows: the BLAST score is multipUed by the percent nucleotide identity and the product is divided by (5 times the length of the shorter of the two sequences). The BLAST score is calculated by assigning a score of +5 for every base that matches in a high-scoring segment pair (HSP), and -4 for every mismatch. Two sequences may share more than one HSP (separated by gaps). If there is more than one HSP, then the pair with the highest BLAST score is used to calculate the product score. The product score represents a balance between fractional overlap and quality in a BLAST aUgnment. For example, a product score of 100 is produced only for 100% identity over the entire length of the shorter of the two sequences being compared. A product score of 70 is produced either by 100% identity and 70% overlap at one end, or by 88% identity and 100% overlap at the other. A product score of 50 is produced either by 100% identity and 50% overlap at one end, or 79% identity and 100% overlap.
VI. Tissue Distribution Profiling
A tissue distribution profile is determined for each template by compiling the cDNA library tissue classifications of its component cDNA sequences. Each component sequence, is derived from a cDNA Ubrary constructed from a human tissue. Each human tissue is classified into one of the following categories: cardiovascular system; connective tissue; digestive system; embryonic structures; endocrine system; exocrine glands; genitaUa, female; genitaUa, male; germ cells; hemic and immune system; liver; musculoskeletal system; nervous system; pancreas; respiratory system; sense organs; skin; stomatognathic system; unclassified/mixed; or urinary tract. Template sequences, component sequences, and cDNA library/tissue information are found in the LIFESEQ GOLD database (Incyte Genomics, Palo Alto C A).
Table 5 shows the tissue distribution profile for the templates of the invention. For each template, the three most frequently observed tissue categories are shown in column 3, along with the percentage of component sequences belonging to each category. Only tissue categories with percentage values of > 10% are shown. A tissue distribution of "widely distributed" in column 3 indicates percentage values of <10% in aU tissue categories.
VII. Transcript Image Analysis
Transcript images are generated as described in Seilhamer et al., "Comparative Gene Transcript Analysis," U.S. Patent Number 5,840,484, incoφorated herein by reference.
VIII. Extension of Polynucleotide Sequences and Isolation of a Full-length cDNA
Oligonucleotide primers designed using a dithp of the Sequence Listing are used to extend the nucleic acid sequence. One primer is synthesized to initiate 5' extension of the template, and the other primer, to initiate 3' extension of the template. The initial primers may be designed using OLIGO 4.06 software (National Biosciences, Inc. (National Biosciences), Plymouth MN), or another appropriate program, to be about 22 to 30 nucleotides in length, to have a GC content of about 50% or more, and to anneal to the target sequence at temperatures of about 68 °C to about 72 °C Any stretch of nucleotides which would result in haiφin structures and primer-primer dimerizations are avoided. Selected human cDNA libraries are used to extend the sequence. If more than one extension is necessary or desired, additional or nested sets of primers are designed.
High fideUty ampUfication is obtained by PCR using methods well known in the art. PCR is performed in 96-well plates using the PTC-200 thermal cycler (MJ Research). The reaction mix 5 contains DNA template, 200 nmol of each primer, reaction buffer containing Mg2+, (NH4)2S04, and β- mercaptoethanol, Taq DNA polymerase (Amersham Pharmacia Biotech), ELONGASE enzyme (Life Technologies), and Pfu DNA polymerase (Stratagene), with the following parameters for primer pair PCI A and PCI B: Step 1 : 94°C, 3 min; Step 2: 94°C, 15 sec; Step 3: 60°C, 1 min; Step 4: 68°C, 2 min; Step 5: Steps 2, 3, and 4 repeated 20 times; Step 6: 68°C, 5 min; Step 7: storage at 4°C In the 0 alternative, the parameters for primer pair T7 and SK+ are as follows: Step 1 : 94°C, 3 min; Step 2: 94°C, 15 sec; Step 3: 57°C, 1 min; Step 4: 68°C, 2 min; Step 5: Steps 2, 3, and 4 repeated 20 times; Step 6: 68 °C, 5 min; Step 7: storage at 4°C
The concentration of DNA in each well is determined by dispensing 100 μl PICOGREEN quantitation reagent (0.25% (v/v); Molecular Probes) dissolved in IX Tris-EDTA (TE) and 0.5 μl of 5 undiluted PCR product into each well of an opaque fluorimeter plate (Corning Incoφorated (Corning), Corning NY), allowing the DNA to bind to the reagent. The plate is scanned in a FLUOROSKAN II (Labsystems Oy) to measure the fluorescence of the sample and to quantify the concentration of DNA A 5 μl to 10 μl ahquot of the reaction mixture is analyzed by electrophoresis on a 1 % agarose mini-gel to determine which reactions are successful in extending the sequence. o The extended nucleotides are desalted and concentrated, transferred to 384-well plates, digested with CviJI cholera virus endonuclease (Molecular Biology Research, Madison WI), and sonicated or sheared prior to reUgation into pUC 18 vector (Amersham Pharmacia Biotech). For shotgun sequencing, the digested nucleotides are separated on low concentration (0.6 to 0.8%) agarose gels, fragments are excised, and agar digested with AGAR ACE (Promega). Extended clones are 5 religated using T4 Ugase (New England Biolabs, Inc., Beverly MA) into pUC 18 vector (Amersham Pharmacia Biotech), treated with Pfu DNA polymerase (Sttatagene) to fill-in restriction site overhangs, and transfected into competent E. coh cells. Transformed cells are selected on antibiotic-containing media, individual colonies are picked and cultured overnight at 37 °C in 384-well plates in LB/2x carbenicilUn Uquid media. o The cells are lysed, and DNA is ampUfied by PCR using Taq DNA polymerase (Amersham
Pharmacia Biotech) and Pfu DNA polymerase (Sttatagene) with the following parameters: Step 1 : 94°C, 3 min; Step 2: 94°C, 15 sec; Step 3: 60°C, 1 min; Step 4: 72°C, 2 min; Step 5: steps 2, 3, and 4 repeated 29 times; Step 6: 72°C, 5 min; Step 7: storage at 4°C DNA is quantified by PICOGREEN reagent (Molecular Probes) as described above. Samples with low DNA recoveries are reampUfied using the same conditions as described above. Samples are diluted with 20% dimethysulfoxide (1 :2, v/v), and sequenced using DYENAMIC energy transfer sequencing primers and the DYENAMIC DIRECT kit (Amersham Pharmacia Biotech) or the ABI PRISM BIGDYE Terminator cycle sequencing ready reaction kit (PE Biosystems). 5 In like manner, the dithp is used to obtain regulatory sequences (promoters, introns, and enhancers) using the procedure above, oUgonucleotides designed for such extension, and an appropriate genomic Ubrary.
IX. Labeling of Probes and Southern Hybridization Analyses o Hybridization probes derived from the dithp of the Sequence Listing are employed for screening cDN , mRNA, or genomic DNA. The labehng of probe nucleotides between 100 and 1000 nucleotides in length is specifically described, but essentially the same procedure may be used with larger cDNA fragments. Probe sequences are labeled at room temperature for 30 minutes using a T4 polynucleotide kinase, γ^P-ATP, and 0.5X One-Phor-Al Plus (Amersham Pharmacia Biotech) 5 buffer and purified using a ProbeQuant G-50 Microcolumn (Amersham Pharmacia Biotech). The probe mixture is diluted to IO7 dpm/μg/ml hybridization buffer and used in a typical membrane-based hybridization analysis.
The DNA is digested with a restriction endonuclease such as Eco RV and is electrophoresed through a 0.7% agarose gel. The DNA fragments are transferred from the agarose to nylon membrane o (NYTRAN Plus, Schleicher & Schuell, Inc., Keene NH) using procedures specified by the manufacturer of the membrane. Prehybridization is carried out for three or more hours at 68 °C, and hybridization is carried out overnight at 68 °C To remove non-specific signals, blots are sequentially washed at room temperature under increasingly stringent conditions, up to 0. lx saUne sodium citrate (SSC) and 0.5% sodium dodecyl sulfate. After the blots are placed in a PHOSPHORIMAGER cassette 5 (Molecular Dynamics) or are exposed to autoradiography film, hybridization patterns of standard and experimental lanes are compared. Essentially the same procedure is employed when screening RNA
X. Chromosome Mapping of dithp
The cDNA sequences which were used to assemble SEQ ID NO: 1-71 are compared with o sequences from the Incyte LIFESEQ database and pubUc domain databases using BLAST and other implementations of the Smith- Waterman algorithm. Sequences from these databases that match SEQ ID NO: 1-71 are assembled into clusters of contiguous and overlapping sequences using assembly algorithms such as PHRAP (Table 6). Radiation hybrid and genetic mapping data available from public resources such as the Stanford Human Genome Center (SHGC), Whitehead Institute for Genome Research (WIGR), and Gέnethon are used to determine if any of the clustered sequences have been previously mapped. Inclusion of a mapped sequence in a cluster will result in the assignment of all sequences of that cluster, including its particular SEQ ID NO:, to that map location. The genetic map locations of SEQ ID NO: 1-71 are described as ranges, or intervals, of human chromosomes. The map 5 position of an interval, in centiMorgans, is measured relative to the terminus of the chromosome's p- arm. (The centiMorgan (cM) is a unit of measurement based on recombination frequencies between chromosomal markers. On average, 1 cM is roughly equivalent to 1 megabase (Mb) of DNA in humans, although this can vary widely due to hot and cold spots of recombination.) The cM distances are based on genetic markers mapped by Genέthon which provide boundaries for radiation hybrid o markers whose sequences were included in each of the clusters.
XI. Microarray Analysis
Probe Preparation from Tissue or Cell Samples
Total RNA is isolated from tissue samples using the guanidinium thiocyanate method and 5 polyA+ RNA is purified using the oligo (dT) cellulose method. Each polyA+ RNA sample is reverse transcribed using MMLV reverse-transcriptase, 0.05 pg/μl oUgo-dT primer (21mer), IX first sfrand buffer, 0.03 units/μl RNase inhibitor, 500 μM dATP, 500 μM dGTP, 500 μM dTTP, 40 μM dCTP, 40 μM dCTP-Cy3 (BDS) or dCTP-Cy5 (Amersham Pharmacia Biotech). The reverse transcription reaction is performed in a 25 ml volume containing 200 ng polyA+ RNA with GEMB RIGHT kits o (Incyte). Specific control polyA+ RNA are synthesized by in vitro transcription from non-coding yeast genomic DNA (W. Lei, unpubhshed). A quantitative controls, the control mRNA at 0.002 ng, 0.02 ng, 0.2 ng, and 2 ng are diluted into reverse ttanscription reaction at ratios of 1 :100,000, 1 :10,000, 1 : 1000, 1 :100 (w/w) to sample mRNA respectively. The control mRNA are diluted into reverse transcription reaction at ratios of 1:3, 3:1, 1:10, 10:1, 1:25, 25:1 (w/w) to sample mRNA differential 5 expression patterns. After incubation at 37° C for 2 hr, each reaction sample (one with Cy3 and another with Cy5 labeUng) is treated with 2.5 ml of 0.5M sodium hydroxide and incubated for 20 minutes at 85° C to the stop the reaction and degrade the RNA. Probes are purified using two successive CHROMA SPIN 30 gel filtration spin columns (CLONTECH Laboratories, Inc. (CLONTECH), Palo Ato CA) and after combining, both reaction samples are ethanol precipitated using 1 ml of glycogen (1 o mg/ml), 60 ml sodium acetate, and 300 ml of 100% ethanol. The probe is then dried to completion using a SpeedVAC (Savant Instruments Inc., Holbrook NY) and resuspended in 14 μl 5X SSC/0.2% SDS.
Microarray Preparation Sequences of the present invention are used to generate array elements. Each array element is ampUfied from bacterial cells containing vectors with cloned cDNA inserts. PCR ampUfication uses primers complementary to the vector sequences flanking the cDNA insert. Array elements are ampUfied in thirty cycles of PCR from an initial quantity of 1-2 ng to a final quantity greater than 5 μg. AmpUfied array elements are then purified using SEPHACRYL-400 (Amersham Pharmacia Biotech). Purified array elements are immobiUzed on polymer-coated glass sUdes. Glass microscope sUdes (Corning) are cleaned by ultrasound in 0.1% SDS and acetone, with extensive distilled water washes between and after treatments. Glass slides are etched in 4% hydrofluoric acid (VWR Scientific Products Corporation (VWR), West Chester, PA), washed extensively in distilled water, and coated with 0.05% aminopropyl silane (Sigma) in 95% ethanol. Coated sUdes are cured in a 110°C oven. Array elements are appUed to the coated glass substrate using a procedure described in US Patent No. 5,807,522, incoφorated herein by reference. 1 μl of the array element DNA, at an average concentration of 100 ng/μl, is loaded into the open capillary printing element by a high-speed robotic apparatus. The apparatus then deposits about 5 nl of array element sample per slide. Microarrays are UV-crosslinked using a STRATALINKER UV-crossUnker (Stratagene).
Microarrays are washed at room temperature once in 0.2% SDS and three times in distilled water. Non-specific binding sites are blocked by incubation of microarrays in 0.2% casein in phosphate buffered saUne (PBS) (Tropix, Inc., Bedford, MA) for 30 minutes at 60° C followed by washes in 0.2% SDS and distilled water as before.
Hybridization
Hybridization reactions contain 9 μl of probe mixture consisting of 0.2 μg each of Cy3 and Cy5 labeled cDNA synthesis products in 5X SSC, 0.2% SDS hybridization buffer. The probe mixture is heated to 65° C for 5 minutes and is aliquoted onto the microarray surface and covered with an 1.8 cm2 coversUp. The arrays are ttansferred to a wateφroof chamber having a cavity just sUghtly larger than a microscope sUde. The chamber is kept at 100% humidity internally by the addition of 140 μl of 5x SSC in a corner of the chamber. The chamber containing the arrays is incubated for about 6.5 hours at 60° C. The arrays are washed for 10 min at 45° C in a first wash buffer (IX SSC, 0.1% SDS), three times for 10 minutes each at 45° C in a second wash buffer (0. IX SSC), and dried.
Detection
Reporter-labeled hybridization complexes are detected with a microscope equipped with an Lnnova 70 mixed gas 10 W laser (Coherent, Inc., Santa Clara CA) capable of generating spectral lines at 488 nm for excitation of Cy3 and at 632 nm for excitation of Cy5. The excitation laser Ught is focused on the array using a 20X microscope objective (Nikon, Inc., Melville NY). The sUde containing the array is placed on a computer-controlled X-Y stage on the microscope and raster- scanned past the objective. The 1.8 cm x 1.8 cm array used in the present example is scanned with a resolution of 20 micrometers. 5 In two separate scans, a mixed gas multiline laser excites the two fluorophores sequentially.
Emitted light is spUt, based on wavelength, into two photomultipUer tube detectors (PMT R1477, Hamamatsu Photonics Systems, Bridgewater NJ) corresponding to the two fluorophores. Appropriate filters positioned between the array and the photomultipUer tubes are used to filter the signals. The emission maxima of the fluorophores used are 565 nm for Cy3 and 650 nm for Cy5. Each array is o typically scanned twice, one scan per fluorophore using the appropriate filters at the laser source, although the apparatus is capable of recording the spectra from both fluorophores simultaneously. The sensitivity of the scans is typically caUbrated using the signal intensity generated by a cDNA control species added to the probe mix at a known concentration. A specific location on the array contains a complementary DNA sequence, allowing the intensity of the signal at that location to 5 be correlated with a weight ratio of hybridizing species of 1 : 100,000. When two probes from different sources (e.g., representing test and control cells), each labeled with a different fluorophore, are hybridized to a single array for the puφose of identifying genes that are differentially expressed, the calibration is done by labehng samples of the calibrating cDNA with the two fluorophores and adding identical amounts of each to the hybridization mixture. o The output of the photomultipUer tube is digitized using a 12-bit RTI-835H analog-to-digital
(A/D) conversion board (Analog Devices, Inc., Norwood, MA) installed in an IBM-compatible PC computer. The digitized data are displayed as an image where the signal intensity is mapped using a Unear 20-color transformation to a pseudocolor scale ranging from blue (low signal) to red (high signal). The data is also analyzed quantitatively. Where two different fluorophores are excited and 5 measured simultaneously, the data are first corrected for optical crosstalk (due to overlapping emission spectra) between the fluorophores using each fluorophore's emission spectrum.
A grid is superimposed over the fluorescence signal image such that the signal from each spot is centered in each element of the grid. The fluorescence signal within each element is then integrated to obtain a numerical value corresponding to the average intensity of the signal. The software used for o signal analysis is the GEMTOOLS gene expression analysis program (Incyte).
XII. Complementary Nucleic Acids
Sequences complementary to the dithp are used to detect, decrease, or inhibit expression of the naturally occurring nucleotide. The use of oUgonucleotides comprising from about 15 to 30 base pairs is typical in the art. However, smaller or larger sequence fragments can also be used. Appropriate oUgonucleotides are designed from the dithp using OLIGO 4.06 software (National Biosciences) or other appropriate programs and are synthesized using methods standard in the art or ordered from a commercial suppUer. To inhibit transcription, a complementary oUgonucleotide is designed from the 5 most unique 5 ' sequence and used to prevent transcription factor binding to the promoter sequence. To inhibit translation, a complementary oligonucleotide is designed to prevent ribosomal binding and processing of the transcript.
XIII. Expression of DITHP o Expression and purification of DITHP is accomplished using bacterial or virus-based expression systems. For expression of DITHP in bacteria, cDNA is subcloned into an appropriate vector containing an antibiotic resistance gene and an inducible promoter that directs high levels of cDNA ttanscription. Examples of such promoters include, but are not Umited to, the trp-lac (tac) hybrid promoter and the T5 or T7 bacteriophage promoter in conjunction with the lac operator 5 regulatory element. Recombinant vectors are transformed into suitable bacterial hosts, e.g.,
BL21(DE3). Antibiotic resistant bacteria express DITHP upon induction with isopropyl beta-D- thiogalactopyranoside (IPTG). Expression of DITHP in eukaryotic cells is achieved by infecting insect or mammaUan cell Unes with recombinant Autographica caUfornica nuclear polyhedrosis virus (AcMNPV), commonly known as baculovirus. The nonessential polyhedrin gene of baculovirus is o replaced with cDNA encoding DITHP by either homologous recombination or bacterial-mediated transposition involving transfer plasmid intermediates. Viral infectivity is maintained and the strong polyhedrin promoter drives high levels of cDNA ttanscription. Recombinant baculovirus is used to infect Spodoptera frugiperda (Sf9) insect cells in most cases, or human hepatocytes, in some cases. Infection of the latter requires additional genetic modifications to baculovirus. (See e.g., Engelhard, 5 supra; and Sandig, supra.)
In most expression systems, DITHP is synthesized as a fusion protein with, e.g., glutathione S- fransferase (GST) or a peptide epitope tag, such as FLAG or 6-His, permitting rapid, single-step, affinity-based purification of recombinant fusion protein from crude cell lysates. GST, a 26-kilodalton enzyme from Schistosoma iaponicum. enables the purification of fusion proteins on immobilized o glutathione under conditions that maintain protein activity and antigenicity (Amersham Pharmacia
Biotech). Following purification, the GST moiety can be proteolytically cleaved from DITHP at specifically engineered sites. FLAG, an 8-amino acid peptide, enables immunoaffinity purification using commercially available monoclonal and polyclonal anti-FLAG antibodies (Eastman Kodak Company, Rochester NY). 6-His, a stretch of six consecutive histidine residues, enables purification on metal-chelate resins (QIAGEN). Methods for protein expression and purification are discussed in Ausubel (1995, supra. Chapters 10 and 16). Purified DITHP obtained by these methods can be used directly in the following activity assay.
5 XIV. Demonstration of DITHP Activity
DITHP activity is demonstrated through a variety of specific assays, some of which are outUned below.
Oxidoreductase activity of DITHP is measured by the increase in extinction coefficient of NAD(P)H coenzyme at 340 nmfor the measurement of oxidation activity, or the decrease in extinction l o coefficient of NAD(P)H coenzyme at 340 nmfor the measurement of reduction activity (Dalziel, K.
(1963) J. Biol. Chem. 238:2850-2858). One of three substrates may be used: An-βGal, biocytidine, or ubiquinone- 10. The respective subunits of the enzyme reaction, for example, cytochtome crb oxidoreductase and cytochrome c, are reconstituted. The reaction mixture contains a) 1-2 mg/ml DITHP; and b) 15 mM substrate, 2.4 mM NAD(P)+ in 0.1 M phosphate buffer, pH 7.1 (oxidation
15 reaction), or 2.0 mM NAD(P)H, in 0.1 M N^HT^ buffer, pH 7.4 ( reduction reaction); in a total volume of 0.1 ml. Changes in absorbance at 340 nm (A340) are measured at 23.5 ° C using a recording spectrophotometer (Shimadzu Scientific Instruments, Inc., Pleasanton CA). The amount of NAD(P)H is stoichiomettically equivalent to the amount of substrate initially present, and the change in A340 is a direct measure of the amount of NAD(P)H produced; ΔA340 = 6620[NADH]. Oxidoreductase activity
20 of DITHP activity is proportional to the amount of NAD(P)H present in the assay.
Transferase activity of DITHP is measured through assays such as a methyl transferase assay in which the transfer of radiolabeled methyl groups between a donor substrate and an acceptor substrate is measured (Bokar, J.A. et al. (1994) J. Biol. Chem. 269:17697-17704). Reaction mixtures (50 μl final volume) contain 15 mM HEPES, pH 7.9, 1.5 mM MgCl 2, 10 mM dithiothreitol, 3%
25 polyvinylalcohol, 1.5 μCi [methyl-3H] AdoMet (0.375 μM AdoMet) (DuPont-NEN), 0.6 μg DITHP, and acceptor substrate (0.4 μg [35S]RNA or 6-mercaptopurine (6-MP) to 1 mM final concentration). Reaction mixtures are incubated at 30°C for 30 minutes, then 65 °C for 5 minutes. The products are separated by chromatography or electtophoresis and the level of methyl transferase activity is determined by quantification of methyl-3U recovery.
3 o DITHP hydrolase activity is measured by the hydrolysis of appropriate synthetic peptide substrates conjugated with various chromogenic molecules in which the degree of hydrolysis is quantified by spectre-photometric (or fluoromettic) absoφtion of the released chromophore. (Beynon, RJ. and J.S. Bond (1994) Proteolytic Enzymes: A Practical Approach, Oxford University Press, New York NY, pp. 25-55) Peptide substrates are designed according to the category of protease activity as endopeptidase (serine, cysteine, aspartic proteases), animopeptidase (leucine aminopeptidase), or carboxypeptidase (Carboxypeptidase A and B, procollagen C-proteinase).
DITHP isomerase activity such as peptidyl prolyl cis/trans isomerase activity can be assayed by an enzyme assay described by Rahfeld, J.U., et al. (1994) (FEBS Lett. 352: 180-184). The assay is performed at 10°C in 35 mM HEPES buffer, pH 7.8, containing chymotrypsin (0.5 mg/ml) and DITHP at a variety of concentrations. Under these assay conditions, the substtate, Sue- a-Xaa-Pro- Phe-4-NA, is in equilibrium with respect to the prolyl bond, with 80-95% in trans and 5-20% in cis conformation. An aliquot (2 ul) of the substrate dissolved in dimethyl sulfoxide (10 mg/ml) is added to the reaction mixture described above. Only the cis isomer of the substrate is a substrate for cleavage by chymotrypsin. Thus, as the substrate is isomerized by DITHP, the product is cleaved by chymotrypsin to produce 4-nitroaniUde, which is detected by it's absorbance at 390 nm. 4- Nitroanilide appears in a time-dependent and a DITHP concentration-dependent manner.
An assay for DITHP activity associated with growth and development measures cell proUferation as the amount of newly initiated DNA synthesis in Swiss mouse 3T3 cells. A plasmid containing polynucleotides encoding DITHP is transfected into quiescent 3T3 cultured cells using methods well known in the art. The transiently transfected cells are then incubated in the presence of [3H]thymidine, a radioactive DNA precursor. Where appUcable, varying amounts of DITHP ligand are added to the transfected cells. Incoφoration of [3H]thymidine into acid-precipitable DNA is measured over an appropriate time interval, and the amount incoφorated is directly proportional to the amount of newly synthesized DNA.
Growth factor activity of DITHP is measured by the stimulation of DNA synthesis in Swiss mouse 3T3 cells (McKay, I. and I. Leigh, eds. (1993) Growth Factors: A Practical Approach, Oxford University Press, New York NY). Initiation of DNA synthesis indicates the cells' entty into the mitotic cycle and their commitment to undergo later division. 3T3 cells are competent to respond to most growth factors, not only those that are mitogenic, but also those that are involved in embryonic induction. This competence is possible because the in vivo specificity demonstrated by some growth factors is not necessarily inherent but is determined by the responding tissue. In this assay, varying amounts of DITHP are added to quiescent 3T3 cultured cells in the presence of [3H]thymidine, a radioactive DNA precursor. DIIHP for this assay can be obtained by recombinant means or from biochemical preparations. Incoφoration of [3H]thymidine into acid-precipitable DNA is measured over an appropriate time interval, and the amount incoφorated is directly proportional to the amount of newly synthesized DNA. A Unear dose-response curve over at least a hundred-fold DITHP concentration range is indicative of growth factor activity. One unit of activity per milUUter is defined as the concentration of DITHP producing a 50% response level, where 100% represents maximal incorporation of [3H]thymidine into acid-precipitable DNA.
Aternatively, an assay for cytokine activity of DITHP measures the proliferation of leukocytes. In this assay, the amount of tritiated thymidine incoφorated into newly synthesized DNA 5 is used to estimate proliferative activity. Varying amounts of DITHP are added to cultured leukocytes, such as granulocytes, monocytes, or lymphocytes, in the presence of [3H]thymidine, a radioactive DNA precursor. DITHP for this assay can be obtained by recombinant means or from biochemical preparations. Incoφoration of [3H]thymidine into acid-precipitable DNA is measured over an appropriate time interval, and the amount incoφorated is directly proportional to the amount 0 of newly synthesized DNA. A Unear dose-response curve over at least a hundred-fold DITHP concentration range is indicative of DITHP activity. One unit of activity per milliUter is conventionally defined as the concentration of DITHP producing a 50% response level, where 100% represents maximal incoφoration of [3H]thymidine into acid-precipitable DNA.
An alternative assay for DITHP cytokine activity utilizes a Boyden micro chamber 5 (Neuroprobe, Cabin John MD) to measure leukocyte chemotaxis (Vicari, supra). In this assay, about 105 migratory cells such as macrophages or monocytes are placed in cell culture media in the upper compartment of the chamber. Varying dilutions of DITHP are placed in the lower compartment. The two compartments are separated by a 5 or 8 micron pore polycarbonate filter (Nucleopore, Pleasanton CA). After incubation at 37 °C for 80 to 120 minutes, the filters are fixed in methanol and stained o with appropriate labeling agents. Cells which migrate to the other side of the filter are counted using standard microscopy. The chemotactic index is calculated by dividing the number of migratory cells counted when DITHP is present in the lower compartment by the number of migratory cells counted when only media is present in the lower compartment. The chemotactic index is proportional to the activity of DITHP. 5 Aternatively, cell Unes or tissues transformed with a vector containing dithp can be assayed for
DITHP activity by immunoblotting. Cells are denatured in SDS in the presence of β-mercaptoethanol, nucleic acids removed by ethanol precipitation, and proteins purified by acetone precipitatioa Pellets are resuspended in 20 mM tris buffer at pH 7.5 and incubated with Protein G-Sepharose pre-coated with an antibody specific for DITHP. After washing, the Sepharose beads are boiled in electrophoresis o sample buffer, and the eluted proteins subjected to SDS-PAGE. The SDS-PAGE is ttansferred to a nitrocellulose membrane for immunoblotting, and the DITHP activity is assessed by visuaUzing and quantifying bands on the blot using the antibody specific for DITHP as the primary antibody and 125I- labeled IgG specific for the primary antibody as the secondary antibody. DITHP kinase activity is measured by phosphorylation of a protein substrate using γ-labeled [32P]-ATP and quantitation of the incoφorated radioactivity using a radioisotope counter. DITHP is incubated with the protein substrate, [32P]-ATP, and an appropriate kinase buffer. The [32P] incoφorated into the product is separated from free [32P]-ATP by electrophoresis and the incoφorated [32P] is counted. The amount of [32P] recovered is proportional to the kinase activity of DITHP in the assay. A determination of the specific amino acid residue phosphorylated is made by phosphoamino acid analysis of the hydrolyzed protein.
In the alternative, DITHP activity is measured by the increase in cell proliferation resulting from transformation of a mammalian cell line such as COS7, HeLa or CHO with an eukaryotic expression vector encoding DITHP. Eukaryotic expression vectors are commercially available, and the techniques to introduce them into cells are well known to those skilled in the art. The cells are incubated for 48-72 hours after transformation under conditions appropriate for the cell line to allow expression of DITHP. Phase microscopy is then used to compare the mitotic index of transformed versus control cells. An increase in the mitotic index indicates DITHP activity. In a further alternative, an assay for DITHP signaUng activity is based upon the ability of
GPCR family proteins to modulate G protein-activated second messenger signal transduction pathways (e.g., cAMP; Gaudin, P. et al. (1998) J. Biol. Chem. 273:4990-4996). A plasmid encoding full length DITHP is transfected into a mammalian cell line (e.g., Chinese hamster ovary (CHO) or human embryonic kidney (HEK-293) cell lines) using methods well-known in the art. Transfected cells are grown in 12-weU trays in culture medium for 48 hours, then the culture medium is discarded, and the attached cells are gently washed with PBS. The cells are then incubated in culture medium with or without ligand for 30 minutes, then the medium is removed and cells lysed by treatment with 1 M perchloric acid. The cAMP levels in the lysate are measured by radioimmunoassay using methods well-known in the art. Changes in the levels of cAMP in the lysate from cells exposed to ligand compared to those without ligand are proportional to the amount of DITHP present in the transfected cells.
Aternatively, an assay for DITHP protein phosphatase activity measures the hydrolysis of P- nifrophenyl phosphate (PNPP). DITHP is incubated together with PNPP in HEPES buffer pH 7.5, in the presence of 0.1 % β-mercaptoethanol at 37 °C for 60 mia The reaction is stopped by the addition of 6 ml of 10 N NaOH, and the increase in Ught absorbance of the reaction mixture at 410 nm resulting from the hydrolysis of PNPP is measured using a specfrophotometer. The increase in light absorbance is proportional to the phosphatase activity of DITHP in the assay (Diamond, R.H. et al (1994) Mol Cell Biol 14:3752-3762). An alternative assay measures DITHP-mediated G-protein signaling activity by monitoring the mobilization of Ca** as an indicator of the signal transduction pathway stimulation. (See, e.g. , Grynkievicz, G. et al. (1985) J. Biol. Chem. 260:3440; McColl, S. et al. (1993) J. Immunol. 150:4550-4555; and Aussel, C. et al. (1988) J. Immunol. 140:215-220). The assay requires 5 preloading neutrophils or T cells with a fluorescent dye such as FURA-2 or BCECF (Universal Imaging Coφ, Westchester PA) whose emission characteristics are altered by Ca** binding. When the cells are exposed to one or more activating stimuU artificially (e.g., anti-CD3 antibody ligation of the T cell receptor) or physiologically (e.g., by allogeneic stimulation), Ca** flux takes place. This flux can be observed and quantified by assaying the cells in a fluorometer or fluorescent activated cell o sorter. Measurements of Ca** flux are compared between cells in their normal state and those transfected with DITHP. Increased Ca++ mobilization attributable to increased DITHP concentration is proportional to DITHP activity.
DITHP transport activity is assayed by measuring uptake of labeled substrates into Xenopus laevis oocytes. Oocytes at stages V and VI are injected with DITHP mRNA (10 ng per oocyte) and 5 incubated for 3 days at 18°C in OR2 medium (82.5mM NaCl, 2.5 mM KCl, ImM CaCl2, ImM MgCl2, ImM 5 mM Hepes, 3.8 mM NaOH, 50μg/ml gentamycin, pH 7.8) to allow expression of DITHP protein. Oocytes are then transferred to standard uptake medium (lOOmM NaCl, 2 mM KCl, ImM CaCl2, ImM MgCl2, 10 mM Hepes/Tris pH 7.5). Uptake of various substrates (e.g., amino acids, sugars, drugs, ions, and neurotransmitters) is initiated by adding labeled substrate (e.g. o radiolabeled with 3H, fluorescently labeled with rhodamine, etc.) to the oocytes. After incubating for 30 minutes, uptake is terminated by washing the oocytes three times in Na+-free medium, measuring the incoφorated label, and comparing with controls. DITHP transport activity is proportional to the level of internalized labeled substrate.
DITHP transferase activity is demonstrated by a test for galactosylfransferase activity. This 5 can be determined by measuring the transfer of radiolabeled galactose from UDP-galactose to a
GlcNAc-terminated oUgosaccharide chain (Kolbinger, F. et al. (1998) J. Biol. Chem. 273:58-65). The sample is incubated with 14 μl of assay stock solution (180 mM sodium cacodylate, pH 6.5, 1 mg/ml bovine serum albumin, 0.26 mM UDP-galactose, 2 μl of UDP-[Η]galactose), 1 μl of MnCl2 (500 mM), and 2.5 μl of GlcNAcβO-(CH2)8-C02Me (37 mg/ml in dimethyl sulfoxide) for 60 minutes at 0 37°C The reaction is quenched by the addition of 1 ml of water and loaded on a C18 Sep-Pak cartridge (Waters), and the column is washed twice with 5 ml of water to remove unreacted UDP- [ΗJgalactose. The [Ηjgalactosylated GlcNAcβO-(CH2)8-C02Me remains bound to the column during the water washes and is eluted with 5 ml of methanol. Radioactivity in the eluted material is measured by liquid scintillation counting and is proportional to galactosylttansferase activity in the starting sample.
In the alternative, DITHP induction by heat or toxins may be demonstrated using primary cultures of human fibroblasts or human cell Unes such as CCL-13, HEK293, or HEP G2 (ATCC). To 5 heat induce DITHP expression, aliquots of cells are incubated at 42 °C for 15, 30, or 60 minutes. Control aliquots are incubated at 37 °C for the same time periods. To induce DITHP expression by toxins, aliquots of cells are treated with 100 μM arsenite or 20 mM azetidine-2-carboxyUc acid for 0, 3, 6, or 12 hours. After exposure to heat, arsenite, or the amino acid analogue, samples of the treated cells are harvested and cell lysates prepared for analysis by western blot. Cells are lysed in lysis o buffer containing 1 % Nonidet P-40, 0.15 M NaCl, 50 mM Tris-HCl, 5 mM EDTA, 2 mM
N-ethylmaleimide, 2 mM phenylmefhylsulfonyl fluoride, 1 mg/ml leupeptin, and 1 mg/ml pepstatin. Twenty micrograms of the cell lysate is separated on an 8% SDS-PAGE gel and ttansferred to a membrane. After blocking with 5% nonfat dry milk/phosphate-buffered saline for 1 h, the membrane is incubated overnight at 4°C or at room temperature for 2-4 hours with a 1 : 1000 dilution of 5 anti-DITHP serum in 2% nonfat dry milk/phosphate-buffered saUne. The membrane is then washed and incubated with a 1:1000 dilution of horseradish peroxidase-conjugated goat anti-rabbit IgG in 2% dry milk/phosphate-buffered saline. After washing with 0.1% Tween 20 in phosphate-buffered saline, the DITHP protein is detected and compared to controls using chemiluminescence.
Aternatively, DITHP protease activity is measured by the hydrolysis of appropriate synthetic o peptide substtates conjugated with various chromogenic molecules in which the degree of hydrolysis is quantified by spectrophotometric (or fluorometric) absoφtion of the released chromophore (Beynon, R.J. and J.S. Bond (1994) Proteolytic Enzymes: A Practical Approach, Oxford University Press, New York, NY, pp.25-55). Peptide substrates are designed according to the category of protease activity as endopeptidase (serine, cysteine, aspartic proteases, or metalloproteases), 5 aminopeptidase (leucine aminopeptidase), or carboxypeptidase (carboxypeptidases A and B, procollagen C-proteinase). Commonly used chromogens are 2-naphthylamine, 4-nitroaniline, and furylacrylic acid. Assays are performed at ambient temperature and contain an aUquot of the enzyme and the appropriate substrate in a suitable buffer. Reactions are carried out in an optical cuvette, and the increase/decrease in absorbance of the chromogen released during hydrolysis of the peptide o substrate is measured. The change in absorbance is proportional to the DITHP protease activity in the assay.
In the alternative, an assay for DITHP protease activity takes advantage of fluorescence resonance energy transfer (FRET) that occurs when one donor and one acceptor fluorophore with an appropriate spectral overlap are in close proximity. A flexible peptide tinker containing a cleavage 5 site specific for PRTS is fused between a red-shifted variant (RSGFP4) and a blue variant (BFP5) of Green Fluorescent Protein. This fusion protein has spectral properties that suggest energy transfer is occurring from BFP5 to RSGFP4. When the fusion protein is incubated with DITHP, the substrate is cleaved, and the two fluorescent proteins dissociate. This is accompanied by a marked decrease in energy transfer which is quantified by comparing the emission spectra before and after the addition of DITHP (Mitra, R.D. et al (1996) Gene 173:13-17). This assay can also be performed in Uving cells. In this case the fluorescent substrate protein is expressed constitutively in cells and DITHP is introduced on an inducible vector so that FRET can be monitored in the presence and absence of DITHP (Sagot, I. et al (1999) FEBS Lett. 447:53-57).
A method to determine the nucleic acid binding activity of DITHP involves a polyacrylamide gel mobility-shift assay. In preparation for this assay, DITHP is expressed by transforming a mammalian cell line such as COS7, HeLa or CHO with a eukaryotic expression vector containing DITHP cDNA. The cells are incubated for 48-72 hours after transformation under conditions appropriate for the cell Une to allow expression and accumulation of DITHP. Extracts containing solubilized proteins can be prepared from cells expressing DITHP by methods well known in the art. Portions of the extract containing DITHP are added to [32P]-labeled RNA or DNA. Radioactive nucleic acid can be synthesized in vitro by techniques well known in the art. The mixtures are incubated at 25 °C in the presence of RNase- and DNase-inhibitors under buffered conditions for 5-10 minutes. After incubation, the samples are analyzed by polyacrylamide gel electrophoresis followed by autoradiography. The presence of a band on the autoradiogram indicates the formation of a complex between DITHP and the radioactive transcript. A band of similar mobility will not be present in samples prepared using control extracts prepared from untransformed cells.
In the alternative, a method to determine the methylase activity of a DITHP measures transfer of radiolabeled methyl groups between a donor substrate and an acceptor substrate. Reaction mixtures (50 μl final volume) contain 15 mM HEPES, pH 7.9, 1.5 mM MgCl 2, 10 mM dithiothreitol, 3% polyvinylalcohol, 1.5 μCi [methyl-3H] AdoMet (0.375 μM AdoMet) (DuPont-NEN), 0.6 μg DITHP, and acceptor substrate (e.g., 0.4 μg [35S]RNA, or 6-mercaptopurine (6-MP) to 1 mM final concentration). Reaction mixtures are incubated at 30 °C for 30 minutes, then 65 °C for 5 minutes. Analysis of [Attet/ιy/-3H]RNA is as follows: 1) 50 μl of 2 x loading buffer (20 mM Tris-HCl, pH 7.6, 1 M LiCl, 1 mM EDTA, 1% sodium dodecyl sulphate (SDS)) and 50 μl oligo d(T)-cellulose (10 mg/ml in 1 x loading buffer) are added to the reaction mixture, and incubated at ambient temperature with shaking for 30 minutes. 2) Reaction mixtures are transferred to a 96-well filtration plate attached to a vacuum apparatus. 3) Each sample is washed sequentially with three 2.4 ml aUquots of 1 x oligo d(T) loading buffer containing 0.5% SDS, 0.1% SDS, or no SDS. and 4) RNA is eluted with 300 ul of water into a 96-well cotiection plate, transferred to scintillation vials containing liquid scintillant, and radioactivity determined. Analysis of [metΛy/-3H]6-MP is as follows: 1) 500 μl 0.5 M borate buffer, pH 10.0, and then 2.5 ml of 20% (v/v) isoamyl alcohol in toluene are added to the reaction mixtures. 2) The samples mixed by vigorous vortexing for ten seconds. 3) After centrifugation at 700g for 10 minutes, 1.5 ml of the organic phase is ttansferred to scintillation vials containing 0.5 ml absolute ethanol and Uquid scintillant, and radioactivity determined, and 4) Results are corrected for the 5 extraction of 6-MP into the organic phase (approximately 41 %).
An assay for adhesion activity of DITHP measures the disruption of cytoskeletal filament networks upon overexpression of DITHP in cultured cell Unes (Rezniczek, G.A. et al. (1998) J. Cell Biol. 141 :209-225). cDNA encoding DITHP is subcloned into a mammalian expression vector that drives high levels of cDNA expression. This construct is transfected into cultured cells, such as rat 0 kangaroo PtK2 or rat bladder carcinoma 804G cells. Actin filaments and intermediate filaments such as keratin and vimentin are visuahzed by immunofluorescence microscopy using antibodies and techniques well known in the art. The configuration and abundance of cytoskeletal filaments can be assessed and quantified using confocal imaging techniques. In particular, the bundling and collapse of cytoskeletal filament networks is indicative of DITHP adhesion activity. 5 Aternatively, an assay for DITHP activity measures the expression of DITHP on the cell surface. cDNA encoding DITHP is transfected into a non-leukocytic cell Une. Cell surface proteins are labeled with biotin (de la Fuente, M.A. et al. (1997) Blood 90:2398-2405). Immunoprecipitations are performed using DITHP-specific antibodies, and immunoprecipitated samples are analyzed using SDS-PAGE and immunoblotting techniques. The ratio of labeled immunoprecipitant to unlabeled o immunoprecipitant is proportional to the amount of DITHP expressed on the cell surface.
Aternatively, an assay for DITHP activity measures the amount of cell aggregation induced by overexpression of DITHP. In this assay, cultured cells such as NIH3T3 are transfected with cDNA encoding DITHP contained within a suitable mammalian expression vector under control of a strong promoter. Cotransfection with cDNA encoding a fluorescent marker protein, such as Green 5 Fluorescent Protein (CLONTECH), is useful for identifying stable fransfectants. The amount of cell agglutination, or clumping, associated with transfected cells is compared with that associated with untransfected cells. The amount of cell agglutination is a direct measure of DITHP activity.
DITHP may recognize and precipitate antigen from serum. This activity can be measured by the quantitative precipitin reaction (Golub, E.S. et al. (1987) Immunology: A Synthesis, Sinauer o Asociates, Sunderland MA, pages 113-115). DITHP is isotopically labeled using methods known in the art. Various serum concentrations are added to constant amounts of labeled DITHP. DITHP- antigen complexes precipitate out of solution and are collected by centrifugation. The amount of precipitable DITHP-antigen complex is proportional to the amount of radioisotope detected in the precipitate. The amount of precipitable DITHP-antigen complex is plotted against the serum concentration. For various serum concentrations, a characteristic precipitation curve is obtained, in which the amount of precipitable DITHP-antigen complex initially increases proportionately with increasing serum concenttation, peaks at the equivalence point, and then decreases proportionately with further increases in serum concenttation. Thus, the amount of precipitable DITHP-antigen complex is 5 a measure of DITHP activity which is characterized by sensitivity to both Umiting and excess quantities of antigen.
A microtubule motiUty assay for DITHP measures motor protein activity. In this assay, recombinant DITHP is immobihzed onto a glass slide or similar substrate. Taxol-stabilized bovine brain microtubules (commercially available) in a solution containing ATP and cytosoUc extract are 0 perfused onto the sUde. Movement of microtubules as driven by DITHP motor activity can be visualized and quantified using video-enhanced light microscopy and image analysis techniques. DITHP motor protein activity is directly proportional to the frequency and velocity of microtubule movement.
Aternatively, an assay for DITHP measures the formation of protein filaments in vitro. A 5 solution of DITHP at a concenfration greater than the "critical concentration" for polymer assembly is applied to carbon-coated grids. Appropriate nucleation sites may be supplied in the solution. The grids are negative stained with 0.7% (w/v) aqueous uranyl acetate and examined by electron microscopy. The appearance of filaments of approximately 25 nm (microtubules), 8 nm (actin), or 10 nm (intermediate filaments) is a demonstration of protein activity. o DITHP electron transfer activity is demonstrated by oxidation or reduction of NADP.
Substrates such as An-βGal, biocytidine, or ubiquinone- 10 may be used. The reaction mixture contains 1-2 mg/ml HORP, 15 mM substtate, and 2.4 mM NAD(P)+ in 0.1 M phosphate buffer, pH 7.1 (oxidation reaction), or 2.0 mM NAD(P)H, in 0.1 M Na2HP04 buffer, pH 7.4 (reduction reaction); in a total volume of 0.1 ml. FAD may be included with NAD, according to methods well known in the 5 art. Changes in absorbance are measured using a recording spectre-photometer. The amount of
NAD(P)H is stoichiometrically equivalent to the amount of substtate initially present, and the change in A340 is a direct measure of the amount of NAD(P)H produced; ΔA340 = 6620[NADH]. DITHP activity is proportional to the amount of NAD(P)H present in the assay. The increase in extinction coefficient of NAD(P)H coenzyme at 340 nm is a measure of oxidation activity, or the decrease in extinction o coefficient of NAD(P)H coenzyme at 340 nm is a measure of reduction activity (Dalziel, K. (1963) J.
Biol. Chem. 238:2850-2858).
DITHP transcription factor activity is measured by its ability to stimulate ttanscription of a reporter gene (Liu, H.Y. et al. (1997) EMBO J. 16:5289-5298). The assay entails the use of a well characterized reporter gene construct, LexA^-LacZ, that consists of LexA DNA transcriptional control elements (LexAop) fused to sequences encoding the E. coli LacZ enzyme. The methods for constructing and expressing fusion genes, introducing them into cells, and measuring LacZ enzyme activity, are well known to those skilled in the art. Sequences encoding DITHP are cloned into a plasmid that directs the synthesis of a fusion protein, LexA-DITHP, consisting of DITHP and a DNA binding domain derived 5 from the LexA transcription factor. The resulting plasmid, encoding a LexA-DITHP fusion protein, is introduced into yeast cells along with a plasmid containing the LexA^-LacZ reporter gene. The amount of LacZ enzyme activity associated with LexA-DITHP transfected cells, relative to control cells, is proportional to the amount of transcription stimulated by the DITHP.
Chromatin activity of DITHP is demonstrated by measuring sensitivity to DNase I (Dawson, 0 B.A. et al. (1989) J. Biol. Chem. 264:12830-12837). Samples are treated with DNase I, followed by insertion of a cleavable biotinylated nucleotide analog, 5-[(N-biotinamido)hexanoamido-ethyl-l,3- tWopropionyl-3-aminoallyl]-2'-deoxyuridine 5 '-triphosphate using nick-repair techniques well known to those skilled in the art. Following purification and digestion with EcoRI restriction endonuclease, biotinylated sequences are affinity isolated by sequential binding to streptavidin and biotincellulose. 5 Another specific assay demonstrates the ion conductance capacity of DITHP using an electrophysiological assay. DITHP is expressed by transforming a mammalian cell line such as COS7, HeLa or CHO with a eukaryotic expression vector encoding DITHP. Eukaryotic expression vectors are commercially available, and the techniques to introduce them into cells are well known to those skilled in the art. A small amount of a second plasmid, which expresses any one of a number of o marker genes such as β-galactosidase, is co-transformed into the cells in order to allow rapid identification of those cells which have taken up and expressed the foreign DNA. The cells are incubated for 48-72 hours after transformation under conditions appropriate for the cell line to allow expression and accumulation of DITHP and β-galactosidase. Transformed cells expressing β- galactosidase are stained blue when a suitable colorimetric substrate is added to the culture media 5 under conditions that are well known in the art. Stained cells are tested for differences in membrane conductance due to various ions by electrophysiological techniques that are well known in the art. Unttansformed cells, and/or cells transformed with either vector sequences alone or β-galactosidase sequences alone, are used as controls and tested in parallel. The contribution of DITHP to cation or anion conductance can be shown by incubating the cells using antibodies specific for either DITHP. o The respective antibodies will bind to the extracellular side of DITHP, thereby blocking the pore in the ion channel, and the associated conductance.
XV. Functional Assays DITHP function is assessed by expressing dithp at physiologically elevated levels in mammaUan cell culture systems. cDNA is subcloned into a mammaUan expression vector containing a strong promoter that drives high levels of cDNA expression. Vectors of choice include pCMV SPORT (Life Technologies) and pCR3.1 (Invitrogen Coφoration, Carlsbad CA), both of which contain the 5 cytomegalovirus promoter. 5-10 μg of recombinant vector are transiently ttansfected into a human cell Une, preferably of endothelial or hematopoietic origin, using either Uposome formulations or electtoporation. 1-2 μg of an additional plasmid containing sequences encoding a marker protein are co-transfected.
Expression of a marker protein provides a means to distinguish transfected cells from o nonfransfected cells and is a reliable predictor of cDN A expression from the recombinant vector.
Marker proteins of choice include, e.g., Green Fluorescent Protein (GFP; CLONTECH), CD64, or a CD64-GFP fusion protein. Flow cytometry (FCM), an automated laser optics-based technique, is used to identify ttansfected cells expressing GFP or CD64-GFP and to evaluate the apoptotic state of the cells and other cellular properties. 5 FCM detects and quantifies the uptake of fluorescent molecules that diagnose events preceding or coincident with cell death. These events include changes in nuclear DNA content as measured by staining of DNA with propidium iodide; changes in cell size and granularity as measured by forward Ught scatter and 90 degree side Ught scatter; down-regulation of DNA synthesis as measured by decrease in bromodeoxyuridine uptake; alterations in expression of cell surface and intracellular o proteins as measured by reactivity with specific antibodies; and alterations in plasma membrane composition as measured by the binding of fluorescein-conjugated Annexin V protein to the cell surface. Methods in flow cytometty are discussed in Ormerod, M. G. (1994) Flow Cytometry, Oxford, New York NY.
The influence of DITHP on gene expression can be assessed using highly purified populations 5 of cells transfected with sequences encoding DITHP and either CD64 or CD64-GFP. CD64 and CD64-GFP are expressed on the surface of transfected cells and bind to conserved regions of human immunoglobuUn G (IgG). Transfected cells are efficiently separated from nonfransfected cells using magnetic beads coated with either human IgG or antibody against CD64 (DYNAL, Inc., Lake Success NY). mRNA can be purified from the cells using methods well known by those of skill in the art. o Expression of mRNA encoding DITHP and other genes of interest can be analyzed by northern analysis or microarray techniques.
XVI. Production of Antibodies DITHP substantially purified using polyacrylamide gel electtophoresis (PAGE; see, e.g., Harrington, M.G. (1990) Methods Enzymol. 182:488-495), or other purification techniques, is used to immunize rabbits and to produce antibodies using standard protocols.
Aternatively, the DITHP amino acid sequence is analyzed using LASERGENE software 5 (DNASTAR) to determine regions of high immunogenicity, and a corresponding peptide is synthesized and used to raise antibodies by means known to those of skill in the art. Methods for selection of appropriate epitopes, such as those near the C-terminus or in hydrophiUc regions are well described in the art. (See, e.g., Ausubel, 1995, supra. Chapter 11.)
Typically, peptides 15 residues in length are synthesized using an ABI 431 A peptide 0 synthesizer (PE Biosystems) using fmoc-chemistty and coupled to KLH (Sigma) by reaction with N- maleimidobenzoyl-N-hydroxysuccinimide ester (MBS) to increase immunogenicity. (See, e.g., Ausubel, supra.) Rabbits are immunized with the peptide- KLH complex in complete Freund's adjuvant. Resulting antisera are tested for antipeptide activity by, for example, binding the peptide to plastic, blocking with 1% BSA, reacting with rabbit antisera, washing, and reacting with radio- 5 iodinated goat anti-rabbit IgG. Antisera with antipeptide activity are tested for anti-DITHP activity using protocols well known in the art, including ELISA, RIA, and immunoblotting.
XVII. Purification of Naturally Occurring DITHP Using Specific Antibodies
Naturally occurring or recombinant DITHP is substantially purified by immunoaffinity o chromatography using antibodies specific for DITHP. An immunoaffinity column is constructed by covalently coupUng anti-DITHP antibody to an activated chromatographic resin, such as CNBr-activated SEPHAROSE (Amersham Pharmacia Biotech). After the coupUng, the resin is blocked and washed according to the manufacturer's instructions.
Media containing DITHP are passed over the immunoaffinity column, and the column is 5 washed under conditions that allow the preferential absorbance of DITHP (e.g., high ionic strength buffers in the presence of detergent). The column is eluted under conditions that disrupt antibody/DITHP binding (e.g., a buffer of pH 2 to pH 3, or a high concentration of a chaotrope, such as urea or thiocyanate ion), and DITHP is collected.
o XVIII. Identification of Molecules Which Interact with DITHP
DITHP, or biologically active fragments thereof, are labeled with 125I Bolton-Hunter reagent. (See, e.g., Bolton, AE. and W.M. Hunter (1973) Biochem. J. 133:529-539.) Candidate molecules previously arrayed in the wells of a multi-well plate are incubated with the labeled DITHP, washed, and any wells with labeled DITHP complex are assayed. Data obtained using different concentrations of DITHP are used to calculate values for the number, affinity, and association of DITHP with the candidate molecules.
Aternatively, molecules interacting with DITHP are analyzed using the yeast two-hybrid system as described in Fields, S. and O. Song (1989) Nature 340:245-246, or using commercially available kits based on the two-hybrid system, such as the MATCHMAKER system (CLONTECH).
DITHP may also be used in the PATHCALLING process (CuraGen Coφ., New Haven CT) which employs the yeast two-hybrid system in a high-throughput manner to determine all interactions between the proteins encoded by two large Ubraries of genes (Nandabalan, K. et al. (2000) U.S. Patent No. 6,057,101).
Al pubUcations and patents mentioned in the above specification are herein incoφorated by reference. Various modifications and variations of the described method and system of the invention will be apparent to those skilled in the art without departing from the scope and spirit of the invention. Athough the invention has been described in connection with specific preferred embodiments, it should be understood that the invention as claimed should not be unduly limited to such specific embodiments. Indeed, various modifications of the above-described modes for carrying out the invention which are obvious to those skilled in the field of molecular biology or related fields are intended to be within the scope of the following claims.
TABLE 1
Probability Annotation
ID NO: Template ID Gl Number Score
1 405310.1. oct g3876615 2.30E-36 Similarity to Yeast D-lactate dehydrogenase (SW:DLD1_YEAST); cDNA EST EMBLC 12235 comes fr this gene; cDNA EST EMBLC12916 comes from this gene; cDNA EST EMBLC10532 comes from thi gene; cDNA EST EMBLC10979 comes from this gene; cDNA EST y
2 48073 l .ό.oct g5669919 3.00E-92 hydroxypyruvate reductase (Homo sapiens)
3 334751.2.dec g262476 1.80E-186 cystathionine gamma-lyase, cystathionase {possibly alternatively spliced} {EC 4.4.1.1} (human, liver, Peptide, 405 aa)
4 237330.8.dec g72 1276 5.00E-70 CGI 0509 gene product (Drosophila melanogaster)
5 053778.1 1 . dec g2905643 1.10E-49 ribitol kinase
6 360645. lO.dec g7022797 0 unnamed protein product (Homo sapiens)
7 334808.1 . ec g2789461 1.20E-244 trehalase
8 997089.7. dec g7023108 0 Homo sapiens cDNA FLJ10830 fis, clone NT2RP4001 143, weakly similar to SUCCINYL-
DIAMINOPIMELATE DESUCCINYLASE (EC 3.5.1.18).
9 237152.1. dec g3355904 3.60E-81 fibroblast growth factor (FGF-18)
10 232851.7.dec g4098959 4.00E-16 tumor necrosis factor receptor-like gene 2 (Homo sapiens)
1 1 083804.1. dec g3851699 4.60E-146 chemokine receptor
12 272721.ό.oct g7330736 6.00E-25 CDK5 activator-binding protein (Rattus norvegicus)
13 461603.4.oct g8250239 4.00E-72 protein phosphatase 4 regulatory subunit 2 (Homo sapiens)
14 332465.2.dec gl931 10 7.00E-250 esk kinase
15 445175.3.dec g35495 2.20E-197 protein kinase C epsilon
16 980541.1. dec g2967685 2.5e-313 serine/threonine protein phosphatase 7 catalytic subunit
17 237996.1. dec g3598974 1.60E-18 protein tyrosine phosphatase TD14
18 243267.9.dec g 1370092 4.00E-20 kinase (Gallus gallus)
19 242082. lO.dec g286105 6.00E-18 zinc finger protein (Mus musculus)
20 019239.1. dec g454158 2.00E-13 zinc finger protein (Mus musculus) 1 899943.1. ec g3869259 1.00E-278 ZNF202 beta 2 443551.1. dec g 1237278 7.20E-105 zinc finger protein 3 897957.1. dec g488557 3.40E-71 zinc finger protein ZNF137 4 90091 1.1. dec g498721 1.80E-21 zinc finger protein 5 999296.1. dec g4469277 6.50E-65 OZF 6 442286.1. dec g5360985 7.80E-12 dJ228H13.3 (Zinc Finger Protein) 7 901978.1. dec g2306773 1.70E-80 zinc finger protein
TABLE 1
Probability Annotation
SEQ ID NO: Template ID Gl Number Score
28 479346.1. dec g5640017 1.00E-72 zinc finger protein ZFP113 (Mus musculus)
29 481750.1. dec gl 020145 1.00E-1 15 DNA binding protein (Homo sapiens)
30 900917.2.dec g 1049301 6.90E-09 KRAB zinc finger protein; Method: conceptual translation supplied by author
31 999415.1. dec g 186632 1.00E-31 Human Kruppel-associated box (KRAB) mRNA, partial eds, clone BRcl744.
32 900680.2.dec g347906 1.20E-93 zinc finger protein
33 902791 ,3.dec g498721 3.40E-1 13 zinc finger protein
34 053826.1. dec g2943716 1.10E-123 25 kDa trypsin inhibitor
35 204932.4.dec g2443870 2.00E-56 R27090.2 (Homo sapiens)
36 400607.19.dec g3142300 8.60E-29 Contains similarity to pre-mRNA processing protein PRP39 gb | L29224 from S. cerevisiae.
37 444248.7.dec g33583 2.10E-33 Ig variable region (VDJ)
38 346599.9.dec g 178848 0 Human apolipoprotein E mRNA, complete eds.
39 480344.2.dec g 190647 1.00E-143 pregnancy-specific beta- 1 -glycoprotein
40 41 1396.24.dec g339943 9.00E-32 Human tropomyosin mRNA, complete eds.
_ 41 302819.4.dec g4589482 4.00E-89 KIAA0925 protein (Homo sapiens)
-o 42 238734.2.dec g6522736 0 dJ777L9.2 (kinesin superfamily protein (KIF)) (Homo sapiens)
43 399525.3.dec g3879121 1.80E-34 predicted using Genefinder; Similarity to Mouse ankyrin (PIR Ace. No. S37771); cDNA EST
EMBLT01923 comes from this gene; cDNA EST EMBLD32335 comes from this gene; cDNA EST
EMBLD32723 comes from this gene; cDNA EST EMBLD33269 comes from thi
44 222795.6.dec g 1657837 7.40E-98 pl lόRip
45 410628.5.dec g3879156 2.30E-97 predicted using Genefinder; Similarity to Mouse ankyrin (PIR Ace. No. S37771); cDNA EST
EMBLT01923 comes from this gene; cDNA EST EMBLD32335 comes from this gene; cDNA EST
EMBL:D32723 comes from this gene; cDNA EST EMBLD33269 comes from thi
46 053649.6.dec g2145122 0 GT334 protein (Homo sapiens)
47 221914.2.dec g849238 4.00E-1 similar to polyposis locus protein 1 (SP:DP1.HUMAN, Q00765)
48 347748.2.dec g7271867 2.00E-22 golgi membrane protein GP73 (Homo sapiens)
49 401482.2.oct g562073 0 Human ribosomal protein L35 mRNA, complete eds.
50 274551.1. oct g36145 2.00E-59 Human mRNA for ribosomal protein SI 2.
51 41 1408.20.dec g571 15 4.10E-45 ribosomal protein L31 (AA 1-125)
52 035973.1. dec g292440 9.00E-85 Human ribosomal protein L37 mRNA, complete eds.
53 456536.1. dec g500654 3.00E-29 Yhrl48wp
54 387807.4.oct g7684537 1.00E-08 similar to KIAA0855; similar to BAA74878 (PID:g4240199) (Homo sapiens)
TABLE 1
Probability Annotation
!EQ ID NO: Template ID Gl Number Score
55 406790.3.dec g495493 2.80E-72 heme A:farnesyltransferase
56 412420.63.dec g3090423 1.80E-15 SA3
57 196623.3.dec g3193336 1.00E-159 DBI-related protein (Homo sapiens)
58 427916.8.dec g5805273 3.00E-14 RNA-binding protein alpha-CPl (Mus musculus)
59 264633.8.dec g3329465 0 NSD1 protein (Mus musculus)
60 337822.4.dec g3342452 3.20E-1 17 PHD finger DNA binding protein isoform 1
61 902943.1. dec g 178281 0 AHNAK nucleoprotein
62 256009.2.dec g 178281 l .όθ-313 AHNAK nucleoprotein
63 231892.12.dec go 164674 1. OOE-07 heterogeneous nuclear ribonucleoprotein, alternate transcript (Homo sapiens)
64 197445.1. oct g 189403 1.40E-13 oxysterol-binding protein
65 348775.1. oct g4165269 3.00E-22 Homo sapiens SYBL1 gene, exons 6-8.
66 336239.5. dec g5441607 3.00E-35 hypothetical protein (Canis familiaris)
67 215660.4.dec g3861217 1.10E-34 UBIQUINONE/MENAQUINONE BIOSYNTHESIS METHLYTRANSFERASE UBIE (ubiE)
391940.2.dec g4191318 0 Human 33 kDa Vamp-associated protein (VAP33) mRNA, complete eds. o~o5 69 978302.3.dec g4530435 2.2e-312 thyroid hormone receptor-associated protein complex component TRAP80
70 228629.1 1. dec g5533375 3.00E-80 cell division control protein 16 (Homo sapiens)
71 01 121 1.5.dec g4567068 2.30E-123 tumor suppressing STF cDNA 4
TABLE 2
SEQ ID NO: Template ID Start Stop Frame Pfam Hit Pfam Description E-valu
2 48073 l .ό.oct 292 537 forward 1 2-Hacid_DH PF00389 D-isomer specific 2-hydroxyacid de 1.80E-3
3 334751.2.dec 194 1234 forward 2 Cys_Met_Meta_PP Cys/Met metabolism PLP-dependent enzyme 3.50E-1
3 334751 ,2.dec 468 1283 forward 3 Cys_Met_Meta_PP Cys/Met metabolism PLP-dependent enzyme 1.80E-0
5 053778.1 1. dec 1452 1655 forward 3 FGGY FGGY family of carbohydrate kinases 2.90E-0
7 334808.1. dec 65 1318 forward 2 Trehalase Trehalase 8.00E-8
7 334808.1. dec 1 17 1676 forward 3 Trehalase Trehalase 1.50E-
8 997089.7. dec 193 1365 forward 1 Peptidase_M20 Peptidase family M20/M25/M40 2.50E-
9 237152.1. dec 215 598 forward 2 FGF Fibroblast growth factor 8.80E-
10 232851 ,7.dec 421 546 forward 1 TNFR_c6 TNFR/NGFR cysteine-rich region 1.30E-1
1 1 083804.1. dec 265 101 1 forward 1 7tm_l 7 transmembrane receptor (rhodopsin family) 9.90E-7
14 332465.2.dec 1654 2454 forward 1 pkinase Eukaryotic protein kinase domain 3.00E-8
15 445175.3.dec 236 51 1 forward 2 C2 C2 domain 4.10E-1
15 445175.3.dec 941 1090 forward 2 DAG_PE-bind Phorbol esters/diacylglycerol binding domain (Cl domain) 1.80E-2
15 445175.3. dec 1436 1651 forward 2 pkinase Eukaryotic protein kinase domain 1.40E-1
- 16 980541.1. dec 817 1824 forward 1 STphosphatase Ser/Thr protein phosphatase 6.80E-
∞ 18 243267.9.dec 209 478 forward 2 bromodomain Bromodomain 1.60E- 19 242082. lO.dec 40 108 forward 1 zf-C2H2 Zinc finger, C2H2 type 5.70E-
20 019239.1. dec 553 621 forward 1 zf-C2H2 Zinc finger, C2H2 type 1 .60E-
20 019239.1. dec 251 319 forward 2 zf-C2H2 Zinc finger, C2H2 type 5.10E-
21 899943.1. dec 1899 1967 forward 3 zf-C2H2 Zinc finger, C2H2 type 2.20E-
22 443551.1. dec 60 128 forward 3 zf-C2H2 Zinc finger, C2H2 type 7.60E-
22 443551.1. dec 284 352 forward 2 zf-C2H2 Zinc finger, C2H2 type 8.20E-
23 897957.1. dec 729 797 forward 3 zf-C2H2 Zinc finger, C2H2 type 5.50E-
24 90091 1.1. dec 129 320 forward 3 KRAB KRAB box 2.00E-
24 90091 1.1. dec 531 599 forward 3 zf-C2H2 Zinc finger, C2H2 type 5.00E-0
25 999296.1. dec 135 203 forward 3 zf-C2H2 Zinc finger, C2H2 type 7.30E-0
26 442286.1. dec 343 528 forward 1 KRAB KRAB box 6.80E-3
27 901978.1. dec 246 434 forward 3 KRAB KRAB box 1.00E-3
27 901978.1. dec 993 1061 forward 3 zf-C2H2 Zinc finger, C2H2 type 5.40E-0
28 479346.1. dec 553 741 forward 1 KRAB KRAB box 1.60E-3
28 479346.1. dec 1399 1467 forward 1 zf-C2H2 Zinc finger, C2H2 type 5.70E-0
29 481750.1. dec 357 545 forward 3 KRAB KRAB box 1.80E-3
TABLE 2
SEQ ID NO: Template ID Start Stop Frame Pfam Hit Pfam Description E-valu
29 481750.1. dec 1371 1439 forward 3 zf-C2H2 Zinc finger, C2H2 type 1.20E-0
30 900917.2.dec 319 459 forward 1 KRAB KRAB box 5.10E-1
31 999415.1. dec 247 396 forward 1 KRAB KRAB box 4.70E-1
32 900680.2.dec 195 383 forward 3 KRAB KRAB box 2.30E-
32 900680.2.dec 783 851 forward 3 zf-C2H2 Zinc finger, C2H2 type 1.10E-
33 902791.3.dec 274 456 forward 1 KRAB KRAB box 1.00E-1
33 902791 ,3.dec 867 935 forward 3 zf-C2H2 Zinc finger, C2H2 type 1 .20E-0
33 902791 ,3.dec 641 709 forward 2 zf-C2H2 Zinc finger, C2H2 type 1.90E-0
34 053826.1. dec 831 1 103 forward 3 SCP SCP-like extracellular protein 1.10E-1
34 053826.1. dec 470 805 forward 2 SOP SCP-like extracellular protein 2.40E-1
35 204932.4.dec 171 617 forward 3 DEAD DEAD/DEAH box helicase 4.00E-3
35 204932.4.dec 718 963 forward 1 helicase_C Helicases conserved C-terminal domain 4.60E-2
37 444248.7.dec 135 386 forward 3 ig Immunoglobulin domain 4.40E-0
38 346599.9. dec 131 499 forward 2 Apolipoprotein Apolipoprotein A1 /A4/E family 6.40E-0
39 480344.2.dec 881 1030 forward 2 ig Immunoglobulin domain 7.20E-0 g 41 302819.4.dec 273 392 forward 3 WD40 WD domain, G-beta repeat 3.50E-0
42 238734.2.dec 97 1233 forward 1 kinesin Kinesin motor domain 6.00E-1
43 399525.3.dec 793 891 forward 1 ank Ank repeat 1.10E-0
43 399525.3.dec 1415 1513 forward 2 ank Ank repeat 1.10E-0
45 410628.5.dec 1447 1545 forward 1 ank Ank repeat 5.00E-1
51 41 1408.20.dec 390 674 forward 3 Ribosomal_L31 e Ribosomal protein L31 e 3.30E-
57 196623.3.dec 143 397 forward 2 ACBP Acyl CoA binding protein 6.10E-
57 196623.3.dec 479 904 forward 2 ECH Enoyl-CoA hydratase/isomerase family 6.20E-
58 427916.8.dec 285 389 forward 3 KH-domain KH domain 7.50E-
59 264633.8.dec 2642 2776 forward 2 PHD PHD-finger 2.00E-
59 264633.8.dec 2780 3010 forward 2 PWWP PWWP domain 1.90E-2
59 264633.8.dec 948 1 178 forward 3 PWWP PWWP domain 3.20E-1
59 264633.8.dec 3317 3709 forward 2 SET SET domain 2.50E-4
63 231892.12.dec 418 612 forward 1 rrm RNA recognition motif, (a.k.a. RRM, RBD, or RNP domain) 4.70E-1
64 197445.1. oct 243 530 forward 3 PH PF00169 PH (pleckstrin homology) domain 7.20E-1
64 197445.1. oct 243 530 forward 3 PH PF00169 PH (pleckstrin homology) domain 7.20E-1
67 215660.4.dec 184 993 forward 1 Ubie_methyltran ubiE/COQ5 methyltransferase family 1.30E-1
CO m
10 i o o- oi oi oi ϋi ^ ω ω ω ω ω w ω ω o
O
cn ro ro ro ro ro — ■ ro — - ro ro Nj — — ro — - — ro ro ro ro ro —■ ro —< ro ro ro — ■ — - ro — ■ — - co
∞ y i o • O O O fc ro o si o o o ro o o o -o o o o so fc ro o n si s,^ o ro^ si o ro n α o n c cn CO s CO co M U -' O' O -' M M 'O M fc co oo cn —■ fc i CO — ' O O — ' I M O Q • fc GiCo Co n ro Cn oo n - .fc. fc fc α ro o- O t -O M b M O J -' fc o fc o ro o o o ro - ro fc i o i — . i-j.
C
C o a ro co ro ro co — ■ ro ro ro — —■ ro —■ — co ro ro co —- ro —■ ro ro ro -' —' r —. co o ∞ O i o o o — • o 4-. sj o —■ —« ro —• o —■ o o -■ o fc W - M O -■ —- ro —' O — • ,- o ^ cng coδ o cn O fc Oδ fcg αS OS sisg Oi O __ o ro -o n ro oo ω co si n o- i cn oo cn ro o n ro oo O_ C_o oo si cn c si cn o O ro α o ω t o SJ ro cn oo ro —' Cn o o -O O o co fc o i ro - c -n oo ro — ' n o o o oTJ
O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O o O O O
^ J
Q Q Q Q O Q Q Q Q Q Q Q Q Q Q Q O Q Q Q Q Q Q Q Q Q Q Q Q Q Q Q Q Q Q Q Q Q Q Q Q Q Q Q ϋ o i Ω u i Ω £ a o α α α αα ααα αααα αα QαQ αdαdddααd αdddd dd ϋ. α α Q. φ ro co ro co co ro ro co co co ro co co — ' co co ro ro ro co ro ro co ro co ro co co — ' o co ro ro ro co ro ro OJ ro
O O
3 O CO CO CO CO CO CO CO CO CO CO CO CO O —1 — 1 CO —1 — 1 CO — 1 O co O co co CΛ co CO CO — 1 (n — 1 cn CO Ω TJ TI TJ TJ TJ TI TJ TJ TI TJ TJ TJ TJ TJ <• <. TJ • , TJ <, T> S S S2 TJ TJ TJ TJ TJ : <: TJ TJ TJ ≤ TJ _k TJ TJ
Ό φ
CO m
(D oJ OJ OJ OJ Oj ro ro ro ro ro ro ro ro ro ro — . _ _ _ _ . _ _ _ __ fc fc fc fc ro o oo sι _ _ _ _ _ _ o o fc fc co co ro ro ro σ o
Ni co co to ω co to — ■ ro ro ro ro CO
Co Co Co o- n o oo o oo co ro o Cn cn n o ro oo i co oo cn c fc cn co fc ro ro oj fc fc i, -
O O CO _ fc O OJ OJ c -n - ro - o — o ro > SI o si o o — ■ si v oo
O fc fc ro , si cn ro OJ CO ro OJ OJ ro Ni ro ro ro
OJ D o C 00 — < • SI D OJ C o o C —O — . • i OJ si ro ro ro ro n co ro oo co si o oo _ CO O n j — cn (> D SI (> o co _ — o Co ro ro — < c» cn cn O0 o o 00 si -o o 00 ro o si oo o 00 OJ O o o o ^ si co C) o Ci o ro ro c_> si o si J 00 O ro o o o ro fc fc o cn O O si - oo O 8 oK oo. ro_ co co fc ro co o o si fc fc Cn r o ∞ co co ro 00 fc o ro o vj — ;
O O O O O o o o o o O O o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o O
^ £ ^ ^ ^ -} ~i -> -) -} -i -} -} -} -} -} -i -i -i -i n -3 -2 -3 -} -} -} -2 -} 3 } > -} -} -2 -2 -} -3 ~3 -} } -π
£ ^ ? ^ ^
Ω O O Ω Ω Ω Ω Ω £ Ω i Ω Ω — ^ Ω Ω Ω Ω Ω Ω Ω Ω Ω Ω Ω Ω Ω Ω Ω Ω Ω Ω Ω Ω Ω Ω Ω Ω Ω Ω Ω Ω Ω Ω Ω Ω Ω Ω Ω Ω Ω Ω Ω
Ω. a Ω Ω Ω Ω α. Ω. Ω. Ω. Ω α ddddddddddddddddddd ddddddddddd dd d d d d d d Φ
— ■ _ ro co _ ro ro co Co _ ro co co ro co ro ro co co ro _ _ ro _ ro ro _ _ ro — ■ ro — • — ■ _ _ — - oj co co ro ro ro ro ro oj o
O
3 co O CO CO I CO CO z 3 co co z H co ; j O co co → → O z O CO = j CO CO → CO co co co co co co co Q. J TJ TJ TJ < _ J TJ ϊ _ J J < Z TJ < =. u TJ TJ S _: TJ < _ TJ TJ < j. u TJ ^ TJ TJ TJ TJ TJ TJ TJ J _
-< —1
"O φ
CO rπ
(0 fc fc fc fc fc fc fc fc fc fc fc fc fc fc fc fc fc fc fc fc fc fc 0J OJ OJ OJ OJ CO CO OJ 00 OJ OJ 00 CO C0 00 CO OJ C0 CO OJ C OJ J oo oo SI sl si sl si si o o o o O cn cn cn co ro — ' O O O O O Oo Oo Oo OO si si si si si O O O fc fc fc fc fc fc fc fc fc o
00 Co fc O OJ O fc O fc fc fc O co — • fc C ro ro co OJ O O O sl O O O O O co o co ro cn o co co co .., _ C - - — C CO - — - — - — OJ -J fc — ' n cn o ro oo o oo o si oo ro o fc o OJ _ fc fc cf ro n I oOn N m) M m N Ki-i ^ CJ- „ Ul n O O si O O X; oo si oo o r ^o C J-O C *O» oo o oo ro o fc cn o ro cn o 00 C oo n ro ro si si si fc oo ro g cn co n fc n o cn cn Ω - o
C co co O n cn o CO o fc cf ro ro ro jv Cn ro cn ro f . n cn co oo oo o o o Sl SI fc _ si S_ co ro fc ro n O fc fc ro Λ i m 'C O O 05 '0 ro ro
O O cn cn co — ■ si OJ ro OJ ro OJ o u uni — i-n Iv. co o ' o c _o> o — oo — o ro m f fc n ft cn ro fc si °j fc oι — *- O oo ro n o fc cn o ^ ^ o _ s " i o ro OJ co _ c i:o ™ > 3 J C 00 fc c n oo oo co o 0 α -' - sj > "O
o o o o o o o o o o o o o oo o o o o o o o o o o o o o o ooo o o O O o o o oo o o o o o o o o
Ω Ω Ω Ω Ω Ω Ω Ω Ω Ω Ω Ω Ω Ω Ω Ω Ω Ω Ω Ω Ω Ω Ω Ω Ω Ω Ω Ω Ω Ωs Ω Ω %Ω ?Ω ?Ω Ω Ω Ω Ω Ω Ω Ω Ω Ω Ω Ω Ω Ω
Ω Ω Ω Ω α Ω α Ω Ω α _._.αααααααα α α Ω Ω α αα αα α α α Ω α α αφ co co — ■ — ■ ^ M - ^ c _ _ _ - w _ ω ^ r r M N N co o ω si ^ ι ι w si ω c c ω co - - ^ M ^ M N ro w Ni M K)
Q O
3 co co co co z CO O CO → z H → CO CO CO CO O O CO CO CO CO co co co co co co co co co → zzj ig co → co → CO O Ω
TJ J J TJ < ζ. TJ TJ TJ 2 < Z 2 TJ J J TJ TJ TJ TJ TJ TJ TJ τJ TJ TJ T TJ T τ> ! > <, _ _ TJ 2 TJ S J J _
Ό φ
_
IΛ C φ
Ω. o
Ul ^ c - 5 α. Q.5 α. _. _. _. _. _. _. _. _. _. _. _. 0- Q. co C α. α. "J <Λ c CΛ C _O. C_ w^ W O. _O. o
H
U E α. o
Q cO CO cO cO cO cO cO cO cO .— O CO CO CO CO — — CM CO CN — - ■— CM CN CO — ^ cθ ^ <N cθ CM '- CM Pi CM c C r r (0 - N - - - - - -
? g b b_ _b b_ _b b_ b_ bPb bPbPb b_ b b bPb bPb b _ b_ b_ b_ _b _b b_ _b b b_ b_ b_ b_ b_ b_ b b b_ b_Pb bPbPb b_ b_ _bPbPb _b b_
"" o o o p p q p o .0.0 o o o o o o o o o p p o p p p p o o o .0.0.0.0.0.0.0.0.0.0.0.0.0.0.0.0 o o o
Q- O- O O OO O- O Ui C CM ^ CO OO C ^ C O CM OO ^ OO CO is O O 00 o oo CO is CO CN CO t — - o is o 'T rs O oO r- oO O OO O OO O- C C C M CM C CM ^' O o o rs co co o ■— (> CO O
CO o T • i—s o LO — O CO — ^ lO O 00 is
UJ CN CN O CN CN CN O CM CN CN O LO 00 00 IS is Ui
Cl O CN — ' CN
CΛ o CN ^r O
CO — - Ui CN — CN CN CM CM CN
CO O
< t fO ^ ^f O M .. C . C . _ C -M . — O 00 O O 00 O O CM O LO rs oo o co
- O O is is is L i O O 00 i O CO Ui CN CO CJ "vf ' O CO - " is O — O O O CN . Ui _ S< Ui ^ 00 i Ui
CO CO CO CO CO O O O CO CO CO O CO CO — - o ^r o O CO is LO CN — OO — - - — i Ui Ui ^ LO ^r o — CO — CM C CM CN CN
o
Q π oo oo oo oo oo oo oo oo — ■ — ■ ^ - - ι- N ^ ^ Λ o- σ ^ σ c σ & α ι> ^ α σ o o o M CN ( Ci o o '5j 't '« ^ ^
90 — ^ ^ ^ ^ ^ t ^ 't Ui Ui LO Ui Ui Ui Ui Ui Ui Ui Ui Ui Ui Ui Ui Ui Ui Ui Ui LO Ui Ui Ui Ui Ui Ui Ui O O O O O O O O O O O O O O O
C a LU H o O o
TABLE 3
SEQ ID NO: Template ID Start Stop Frame Domain Type
64 197445.1. oct 2146 2196 forward 1 TM
64 197445.1. oct 2380 2436 forward 1 TM
64 197445.1. oct 2158 2214 forward 1 TM
65 348775.1. oct 232 288 forward 1 TM
65 348775.1. oct 937 999 forward 1 TM
65 348775.1. oct 253 306 forward 1 SP
65 348775.1. oct 1523 1582 forward 2 TM
65 348775.1. oct 942 1010 forward 3 TM
65 348775.1. oct 989 1048 forward 2 TM
65 348775.1. oct 235 282 forward 1 TM
65 348775.1. oct 967 1017 forward 1 TM
65 348775.1. oct 907 975 forward 1 TM
66 336239.5. ec 1670 1744 forward 2 SP
66 336239.5.dec 1317 1379 forward 3 TM
66 336239.5.dec 2417 2485 forward 2 SP
66 336239.5.dec 1217 1279 forward 2 TM
66 336239.5.dec 2217 2282 forward 3 SP
66 336239.5.dec 1725 1784 forward 3 TM
66 336239.5.dec 221 1 2273 forward 3 TM
66 336239.5.dec 852 935 forward 3 TM
66 336239.5.dec 2226 2276 forward 3 TM
68 391940.2.dec 2125 221 1 forward 1 SP
69 978302.3.dec 2319 2372 forward 3 TM
70 228629.1 1. ec 917 979 forward 2 TM
70 228629.1 1. dec 944 997 forward 2 TM
70 228629.1 1. dec 917 1009 forward 2 SP
71 01 121 1.5.dec 1515 1580 forward 3 SP
71 01 121 1.5.dec 1515 1586 forward 3 SP
71 01 121 1.5.dec 1515 1598 forward 3 SP
71 01 121 1.5.dec 1515 1577 forward 3 SP
ω τ- o oj
Q ϋ o o cj o o o o o ϋ o ϋ o o o o o ϋ ϋ ϋ o o o o o o ϋ o o o o cj ϋ o o o υ o cj o o c o o o o o u o o u u o c o υ υ o ϋ u ϋ ϋ
_ o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o
_i — ddddddddddddddddddddddddddddddddddddddddddddddddddddd dddddddd c" C C C C CO C CO CO C C C O C C C C C C C « O C C CO C CO C CO Cr> CO C CO C C CO C C C C C O
C IΛ LO LO LO LO LO UJ LO LO LO LO LO LO LO LO LO LO LO LO LO LO LO LO LO LO LO LO LO LO LO LO LO LO LO LO LO IO LO LO LO LO u ooooooooooooooooooo ooooooooooooooooooooooooooooooooooooo ooooooo
H ^^ ^'^'^t^ ^'^^^'^'^'^ r '^^^'* ^^^^^'^^^^'*'^'^ ^'^ '* '^'*^^^^^'*^^ '*'*''*^ ^
ω c o — .
O εQ
U B
Q O ϋ ϋ O ϋ ϋ CJ O O ϋ O U ϋ O O O U O ϋ O O O O U O ϋ ϋ ϋ O O U O O U ϋ O O ϋ O O O O O ϋ ϋ tj ϋ O O ϋ O ϋ O O O O J U C) J o
MS i—i o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o
90
_! ddddddddddddddddddddddddddddddddddddddddddddddddddddddd ddddddd c c c co co co co co c c co c co co co co co c c co co c co w co c c co co n co co c co co c co co co c co w co n o S LO LO LO LO L L LO LO L LO L L L L L L lO lO L LO U) LO L L L lO L i_ LO LO LO lO L L I^ t^ o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o t-1 "^ ^ ^ ^ ^ ^ ^ '^ α -tf ^ -tf -tf ^ -tf 'vt -t '* '* -* '* ''* ^ '* '* -* -* -* '* '* '* ^ -^ '* -*
co a Z
LO— — CM CM ■— COO COLO LO f Is Of Ol 0100 CM f f (000( fιnW*nnθ)ι-N**fffθn*
N in n τ- ω - (o ω ω o o w f « oo ^ n w n o ι S τ- w s »iD ω ffl i) ro s » o w N (βc θ - W - τ- s N co w w ιO U) io ω ω ι- (θ n w o w o n n n » N n ffl n o o3 f tD ^ ι» * * s t-ι- c f * f * f N f M- ιn f ω w ι ffl (D !θ(o α) (θ (θ (D (θ (θ (D !D (β <o LO CO — — co — co -— — — co — — co — — LO — — — co — •— rv, ~ t-~ — _ T- ,_ _- - IΠ T- ~ •— •— — — i- — •— — •— — — — — — —> ■— — — •— LO CM CM — CD CO CM CM CO CO CO '
MS IΛ JN S O) f lf) O CO t 00 000000 O100 CO •* CM f 0000 - f (D (D W n t (D f 01 S S MO ffl t f 01 λl f <-
© r —o —- ■*o —« tto —o —o ■—N f O ■—O ■—N-'-N f ^ —' — N o •*-i*-o - >- — f ^ —N —W t wf nt « —« —N —N τ-N_w —n f fnN --N —w —w T- —o —n •—n —o ■■-n,- ■—f—f■—* T-f —* —f •—f■— —ini —n —io f
Ul
H — — — CM — ^ — - — - — . — - —- U X X X X x tr x x x x - — T T X OC X X X ii. — T T — - X _ — CO CD —,
T T — X X X T: x X X X X X I H U- X IS X X X I I I X - X X I D X X α. CO CO 00 C f O O O CM CD X - ± o r X: x— xCO x00 xCO HO O O LO O LO O 00 — —
LO LO CD O f CO — CO — LO s o o f o o l- ^* o CO O LO T o — CD — CO IS f 01 f CO CO
CM 00 O LO i-Nf MOffl ^ o o — - rv σ> 00 CO — — 00 CO 00 — — O CO O CD
CD 00 CO CJ) LO O O CO IO CO — - ooioiNin o f o Ol O Ol f S CAI 0000 LO CO is CO CD O O f CO CM CD C J) ^ o CJ) — - o LO O O f CD o o o w fj o in — r-v co σ CO 0000
Ol LO CO 00 o- ui — r-v i-v — co CM f CO — IS CM — — —
CM f CO -<fr C inJ) C O Os —fCOsl O CM — f m N CN LO CM CO -* OO C
LO — — CM CO is f f — * CO CM CO — LO CM J- CO CM f CM CM CM ϋ O O O O ϋ ϋ o o O CJ o o o o o o O O CJ o o o o o o o o O O O O o o o o o o o o o o o o o o o o o o o q o o o o o o o o cbcDcocococdcocbcbcbcbc
OOOOOOOOOOOOOOOOOOOOOOOOOOOOOOOOOOOOOOOOOOOOOOOOOOOOOOO — — — — — — — — — — — ■
•— — — — — r- •— — i- —- — i- — — — — — — — —— — — — — — — T- — — — — — — T- — •— — — ,- —• ■— — — co co co co co co co co co co co c co co co co c co co co co co co c co co co co co co c co co co co co co co c co co co co oo c co co co co co co c c co co co o co c n
LO tO LO LO L L L L LO LO L LO L LO LO W L LO LO LO LO Ul U) i_ LO LO L L U5 LO LO LO LO L U5 LO LO LO LO IO L L O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O OO OO OO OO OO CO OO OO OO OO OO O
CM CM CM CM CM CM CM CM CM CM CM
_i _) oo s ^ m s ffl oo oo — Lo o o r o co L f co f r-v
CO Ol UJ CO OO CO f CO — - O ffl O 03 N »- (0 σ) ffl lll 0101 f O O N (D W O O ι) 00 S W S » O β N S O S lΛ in θ n n S f O ffl t0 !D n N f - - (0 n N N N 00 00 tD o co (s θ '- t o ^ (θ o) ffl '- o ffl !θ '- * "* t ffl f ιϊ) N (θ s ιn s n o ιo o ( n t o3 f n '- θ) θ θ o o « s o o τ- t o o o o ι- N N ι- 't '- o a) n N N f rv. o _ |v. cD r^ r^ cD co _ r . cD cD o ι-- r rs. rv. L i uι -- ι |s. r^ r^ co ^ D CD co o cD co oo ω
LO — oco LO — f j) L r Lθoo oo oθ cj) |. -— ooco — f o f ifo fioifn of wf iffl ifnof sf of fsNf Nf Of Nf fs afiafjof ofi ofi
X C O C L o C C
_ o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o (*1 o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o
90
O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O o o o o o o o o o o o o o o o o o o o o o o o o o co co co c co co co co o co c c co co co c co co c co co co co co co co co c co co co co co c co co co co co co n
O o U UoLoO LoO LoOoLO UoI UoI UooLOoLO LoO LoOoLO UoUoIoO UoI LoOoLO LoO UoUooU oLO IoOoU oUI UoI UoLoO LoOoUI UoUoI UoI IoOoLOoLO LoO UoIoUI UoI UooU oWoooooooooooooooooooo
O
S CD CO — iΛ w in n in n w M W o n c s ιo
S I
o o υ o o o o o o o o o o o o o oό o o oq oό oό oό oό oό oό oo oq o d d d d d d d d d d d d d d d o d s s s s s s s s s s s s s s s s f f f f f f f f f f f f f f f f f f f f f f f f f f f f f f f f f f f f f f f
_- x a>
H
U OO Ul O OO LO OO CN f LO OO O CD S O Ol LO Ol f CO OO t f S CO f f S OO O O CD f O CN f S C0 00 f 00 C0 O CM O CM _ CD 00 C0 00 O S O 00 f CD O CD f C0 CD O C0 CM CM O O S O C0 CD CD 00 - O CM -r- C0 CM C0 00 f CM CM CΛ C0 O O C0 01 CM C CO _l S CO CM 01 C '<a- CO LO L O CO f S -- f S S S O CM f O CO C i- CM CM f CO CO CM C0 CM CO CD _ Ul Ul S Ul S S S CO S S 00 CO S 0000 S CO f f f CD CM Ul -- f o _ _ f n * _ _ f ιn sN \i N n w r- N _ f in (D s _ - ffl - - τ- - τ- r- ^ . - - - τ- - τ- τ- r-- ^ - - - - ^ τ- - - - - - F τ- w ιn _ ιn o * f f f f Ul CO CD CO O CD CD f C CD OO CO S CD CJ) CD CO C f O LO -<t
01 01 -- S 01 CD -- O S f CO CM S CN CO CO LO LO CO — Ul S -- 0O 00 0O CM CM CM CM CM CM -l CD S 00 Ol C LO S 00 LO CD S CM f S 00 CM CO CD _> S σioi- sioss OO OO CM CO CO 'v- LO CO CD S O CO •- O — — - C CO ^ LO O CM τ- LO LO LO CD CO O O O O O O -- -^ -<- CN CN CM CM CO CO CO f f 'lj- f Ul Ul Ul CD CD LO CD CO CO CD CD CO C ι- ft - ι- oι Cλi w n o m ιo β <c _ to eo _ o> oi τ- - τ- '- τ- - τ- - - τ- - - - - - ^ τ- r- - - - - - ^ - CD '— ^- — ^- CO CD —
11 r: — X I ϊ 11 r: — CM X CO CD — LL l I I I X H CD CD I I — — - I CM CD 01 I— I I I S CO LO I I I I l
O LO l I CD i- CN O CD l I CN 1S 1CO 1CD 1CO 1CM LO ϊUl_UlϊUlϊCOϊCDϊi-oSϊLOϊCD CO OC X I CN — CO O CO CD f Ol CO H I— CO CO CN f f CD 0 CD O O 1- o s CD CD Ol •— 0O CD S — S CO CO CN CO CM OO OO OO OO — CD S S f f Ul CM S f CO S Oi σ O O f f CM CO CM OO OO CM CM CD O OO CO Ul CD CD OO 01 T- * CO 1- S 00 ι- S O C 01 00 CD C0 '-- 00 f S S O L0 O L0 — f LO LO f f S S CO LO Ul O CO CM co s σi σi — - — - σ ui σi s s oo cD s σi s o s CO Ol Ol CM CM O 0 CO f f CD f CO CO f S LO OO OO CO CO S S CM O LO CO CO — CO CO f CM CM S S CO CM CM f CO S U OO — ι- Oi σ 0O CM 0O 0O 00 CD 0O 00 — •<fr T- CD f 00 CO Ol 0 — 0 CD Ol O CM O O LO LO -- CO CO LO CM CN OO O S OO CN L CO CN CM OO _ CO CO CO — O0 τ- CM CO OO UI UJ CO CO UI LO — CM — s r CM CM f S S U> Ul CM CD CO CD LO CM S CO O OO CD CO CO CO — CO CD CD OO S CO OO S LO CD CD CN S CD CD f OO CO CM LO CM — CD CD CD S S CM OO CN CM CD CD CN CD CD "» CM f Ul 0000 S S LO 01 — — — CO CO f LO C - - O OO - f W W r- tD OO COLO LO U Ul CN CO O CM CM f f Ul CO— f f CO CD CD CO COf f Tt -f f M f CD CD CO CO COCM CM CO COLO CM CD CO COLO CO O f
CM CN CM CM CN CM CM CM CN CM CM CM CO CO CO CO CO CO CO CO CO CO CO CO C C CO CO CO CO C C CO CO O CO CO CO CO CO CO CO CO CO C CO n
CO _. I I I I I _. I I- I I I I I I CD I I I I_. C
o o o o o o o o o o o o o o o o o o υ υ o o o o o o o o υ o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ T5 TJ TJ TJ XJ X! XJ XJ TJ TJ TJ TJ TJ XJ X! T! T5 TJ TJ TJ TJ X! XJ T3 TJ TJ TJ TJ TJ XJ X) TJ T3 TJ TJ XJ TJ TJ TJ TJ T3 TJ TJ TJ O
^ ^ r ^ r ^ ^ T ^ r ^ ^ ^ ^ ^ T T -r T T÷ -r τ τ T T T — — T-^ — ^ ,- ^ ^ — ^ — ^ — ^ — ^ — ^ — -^ ^ ,- ,- ,- ,- d d d d d d d d d d d d d d d d d cboόoόoό cbcό oό oό cό oό cό eb cb c cόcbo cό oόco cb rø
IsS SS SS SSS SS SS S SSS SSSS SSSS SS SSSSSS SSS S S S SSS SSSS SSSf fffff f ff fff f fff f
S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S CO CD CO CD CD CD CO CO CD CD CO CD CD CD CD CO CD C
00 S OO O) O — S OO S S LO CO — ∞ LO CD CO S OO O CD CD CD S CO CO O OO S i- CO OO S O O O CO — LO CO CO CO CO -- LO O
OO CM LO CN O CO CO OO CO CD CN LO CD O CD S O LO CD O CM S LO CO O LO O S O — O CO LO O O — — S O O S O i- O O O — OO O CO f Ul f CO Ul OO CO CO Ul CD O CM CM CD S O O CD CM — O CM O CO — CO — CO — 1- CM CO CO CN CN OO CO CM CO CO CO CO CO CO CO CO CO CO CO CO CD OI CD CD LO LO — CD OJ OI CD O CD σ> — 01 01 — — oi — — — — — — — — — 1- — — — _ _ — τ- f- — — — — — — — — — — - - - - - - - - cιι cn o) - τ- - - cr) -
O CN CO S — f CD O O OO CD CO S CO O CO O
O CD OO O -.- CD S 'S- S CO f f LO CM S OO CD CD CO CO CO LO i- — l_l S CM O CD f f LO CO C S S CO CN CD CD O) T- LO _> S CO C CO LO CD CO — CM — f — — LO S f LO LO — — O S LO _ CD O S S CD O CN S S S OO O -- CN CN CN LO LO L CO S S O CO LO CO CD O — CN CM CO CO -tf S S S OO CO CD CD Ol Ol O O O O O O — — 00 S S S — -- OO CD CD S S CO f L CO CO CO f f f Ul CO CO CO CO CD CO S S S S S S S S S S S OO OO OO OO CO CD CD CD CD CD Ol Ol Ol Ol Ol CD CD CD CD Ol Ol — — — — — i- — — — — — S S CO — — — — CD —
T- — — —• ■— — CD — ~ " — — CD i-lr_ii ι l sc f r: I I Ul CM 00 I s X X I c
O S CD l f Si —z I_i CDl f l f l fl — i OrljOCicDιCD_ —ι —ι _1ι0-1_CDiCDιCiN)rCO L_-iCOi —fli — Ln — CO S CO l l 0000 Ul CM l LO f CD CN So —O CcoD ICD I— h—- S -- CD CO 00 — OO CD f OO CM CM CD CD -— — CO CO CD CO CD S S S S S CN — — CO CD O LO CO OO OO CD LO CD Ul Ol Ul CO CD T- — S S CD CM f CO C θ ι- uι -- cM S -- co oo f o <v- σ) 1- CM CM Ul f O O S Ul Ul Ul S LO OO S O Ul OO CO CO O CO — S OO O O OO 00 LO CO CO O f f CO CO C CD f f S f CD CN LO LO CO LO CO CD O f OO OO CD Ol CD CD CO CN CM CM CO CO CO CM S CO S CO OO f f O O CM O CD CO O 00 Ol CO O CD S O 01 CO f 0 CO CN O f CO OO CN CD O S CO OO CN Ul S S S S S S OO f f f LO O CD S CD CO CD S O U OO OO CO CO 1- — CD 1- S OO CD O CD Ul CD f S Ul O CO — 00 f — — O CO LO OO CO CO O O S S U UI CO CO CO OO CO OO CM CM CN CO CM UI CO C f Ul CO S S S CN f Ul CO CO CM CO CD CM CM CM Ul CM CM CD COCO CD CM CM UI CD CO CM CD CO CO CO OI OI LO LO — — co — — — co co— CD CO— COCM Ol CO CO f — — CD CD CO CM COCO Ul Ul CO CO COCO Ul CO C
0 0 0 0 0 0 0 0 0 0 0 0 0
0 0 0 o o 0 0 0 o o o o o o 0 0 0 o o o o υ υ o o o o O Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ cι> φ φ φ φ φ φ φ φ α> φ φ φ φ φ φ φ φ φ φ φ φ φ Φ Φ φ φ φ φ α> Φ © X Ό X Ό Ό Ό Ό Ό Ό Ό Ό Ό X o _ XJ XI XJ XI X) TJ TJ TJ XI XI TJ J Tl TJ TJ J XJ XI XI XJ XI XI J XJ XI J XI XJ XJ _. _ ,_ ,_ _. _ _. _. _. _. _. ,_ ,_ , 00 CO 00 0000000000 oo oo αo oo αo oo co oo oo oo oo oo oo oo oo oo oo oo oo oό cό — — — — — — —. — — — — , — . odd o o 000 o o o o o o 000 o o o o o o o o o o ddoόc oόoό oόcboόoόoόoόcόoόoόc
CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO O CO CO CO CO CO cocossssss sssssss CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO m co ro ro cocossssss sssssss o s s s s s s s s s s s s s s s s s s s s s s s s s s s s s s s s s s s s s s s s s s sscocococococococococococococ o CO CO O O CO CO CO CO s s CO CO CO s s s
CO CO CsO CO CsO CO CO CO CO s s
CO CO CO CO CO CO CO CO CO CO ro CM CM CM CM CM CM CN CN CM CM CM CM CM CM CM CM CN CM CM CM CN CM CM CM CM CM CM
f ff f f f ff ff f ff fff f -tffff ffffffff ffff ff fffffff fffffff ff
C» CD „ CB C- ∞ s| vj s' sl sl si s| sl s! s! sl s! sl s' sl sl sl si si sl cD CD CD CD
co co co co co co co co co co co co co co co co ω co co co co co co co ω ω co co co co co ω co co co co ω c co co co co co co co co ω co — _ o — co co ω ω ω ω co co co co co ω ω ω c ω ω ω ω ω ω cD CD σi oi σi σi σi cΛ cD CD CD CD σ OT l l l | -v| vs| ^ ^_ 4_ j_ jv. jv. ^ ^v jv. ^ 42. :i. _ j_ ^_ j_ ^ ^. j_ ^_ o O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O ffl 03 _ 03 _ _ C ffl _ CO OO ffl „ CO OD _ CO CO β _ CB 0101 C0 01 0101
CD CD C0 00 CD C0 O O O O O O O O O O O O O O O O O O O O 4- . . 4- 4- 4- 4- 4- 4- 4- 4- 4- 4- 4- 4- 4- 4- 4- 4- 4- 4- 4- <D — CO CO — — CX3 CD C» CD Ca C» C» CX) C» CX) CX) C» C» C» C» C» ra si si si si si b. x i_.i_.i_. Q. Q. Q. Q. ix ix _______αd______00PPPPPP0000PPPPPPPPPPP0PPPPPPP0PPPPPPP?P
Φ CD Φ φ φ φ φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ _ tt Q. _ _- _- _ Q_ _^ _ _ _ _ _. _. _ _ _ _ _ _ _. _. _. _. _ _ o o O O O O o o o O O O O O O O O O O O O O O O O O Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ
O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O
4- O J- — CO ro ro 4-ια — -co — CD 4- ro 4- roco 01 co ci ->- σ 4- _ roN *» ω ω ω co ro ro ro coco cn -cα rocα » r co co co 4- oo co 4- 4- cn ro rocα cn co σι co 4- 4- co co 4- si
CO i CO 01 CO sl M CD IO CD cD co cD O CD -'- o ro cn — cn co 014- — 010 * 010 -J- ro co σ 01 si co co si co si si — cπ o oo si *». — • cn cn co co 4- — —. OJ cD co co .t- _ sj co ro _. cn
CO si CO — M ro 4- j i ω co oo co σi — OO CΠ J- — -*• CO 4- o CO si CO CD — O CO — 00.1- O o cn ^ ro ω ro — ^ co co σi si * oo _ ooω ωθ M Co c -' ro ji Co ω ro M _ (θ ^ o
CD ro oo co oi 00 O — 00 CO C04- O 00 si rO C0 O sl C0 CD CD — sl CO O CO CO GO W — CD CD sl sj M Oi ro O CO C -OO OKO O co 4- OO OJ CD — sj -i cn cn co ro co — co o co cD ro -- oi ω oi n ro oo AϋlO to CO 03 CO O C04- -'- CO Cn OO Cn 4- CD — CD CD O Ol si si IO 00 — cn 0000 CD 00 .&. CD 00 CD CO CO CD Ol s| co — o co — co .ι ro cD CD Co cn cn - co co cιι cπ 4- si co σ> — cn o ro cn co co - -J si si — ro σι σι ->- co co 00 co — 00 ro cn N> O sl -"■ 00 si o) cn cn cn si cn co cD Co cD oo co ω o co - - U O O | l M 01 Ul ^ - <i O) -Α - O O! CO
CO CO J- CD O _ ro 4- CD o o-O)θ4J noMr Φθoιωoo IS. Is) 00 CO — 4- 01 W 0000 CO 00 -l- -'- G00000 — CO IVJ CO 0) _ N T T t> M O 0000 U1 - A „ --I I U O C000 T
X I X X I 13x01 ~Ooιιιχ-χ lιιι -χ XX II lsiχχχ
-Ol- —.. O) O) - _ - 1-_ ^ro ^— I— noi _. _ Ξ — — Ξ Ξ Ξ Ξ oo _ _ _ _ "" ϊ ϊ ϊ Ξ -
_k _ _i. _v _ _ θi 4_ — — — — - - co co oi oi - - -' — - - - - vi o oi Λ ω ω co ro M - - β (D . _ . _ _ -j (o - -' -' -' -' -' -'-' -' -' -* - -' 4- 4- o ro o r M M r rύ oi o) oι w ^ o ω ^ -' _ 4i ω M - - - o o -' r _ o ω ωc _ _ _ oo cn oi 4» M _ - σ) -' i\3 θo o o o M r M - _ - M - si cn 00 ω cD oo cD oi oi o oi oi co — σ si σi 4- oo co oo si si co o — co ro ro o ->- o ιo co ro co o co -». oo o o o o si σι cn oo co co co ro 00 cn M O)
4- σι oo cD σι σι — — co co A A O UI O -Ό <O o c r o o co io oi c ω -' u ω
_ - _ _ - _ ω ω -' -' - _ _ - -» -' - r - - _ - - - - - - » - - - - -' r ω -' - -' -» -' -' -» -> -' - -' -' - -' -» -' -' -' -' -' -' -' i ιy _ 3 - -' M M C * ω n cn ^w ji ω - Φ M ω ω r M o o o Φ - Oi cn ^ ω - ω ii -' ii -' -' Oi - ^ ffl roN N -' M -' M O - ^
CD 4- I si ω — CO CO OI CD CD CD — O Cn sl CD CD CD O 004- CD CD CD — IO CO O M M Ct) CI1 Cri ∞ Φ ^ M M Φ _ ω θ O M ffl ω _ » _ ffl W O Cn - 4^ i» Ui ω θθ rθ - O O si r Nl σi si c σi si oo cn oo ro cn si co oo 4- co σι co 4- o cn — — — oo co σ co σ ro oo co co cn ->- σι o cn co ro o cn — cn oo cn o
OO OO ∞ ∞ C- CD ∞ ∞ CXI ∞ CB ∞ ∞ ∞ CS CB ∞ ∞
a.
_ cn o ω _CQ -na — ro roco _ coca ca ca _ — ca cn cn ^ cn _ si ca 4- cn *- ro coca _ 4-ca cn _ co ro _ _ cn co ro ro 00 σi 01 cn cn 4- ro ro ro 0 co cn 4- cn ro cn — - 00 ro 4- 01 — cn cn cn
O si — si 4- ro —- cn — si si O in ω ω u cn cn 0 — ^ — - ro co cn ro ro co co ro ω co ro si ro 010000 0 ro ro co ro — co cn ro o) ro — cn cn ro
4- ro cn cn ω co ___ co cn cn — ω —- cn ro ro σi ro *-. ro ro — - — —
4- ro ro cn — co cn co ro ro 0 cn cn — - ro cn co I Ti ro 0 1 — cn ro 4- ro 01 ro cn co *-. *. cn 0 ω -_ X σi ω CO I I I I cn I si X * I.
I - III _. X co I I I III X I I I O sj co X I I I I I ± 00 III I I X X I I I X X I I X X I TJ
- — — * -" k - * -^ ~ - k - * — * si si . . . . ro 00 00 σ> cn cn 4- 4- 4- 4- n cn 10 10 ro ro ro ro ω ro — -. — - r cn ro cn ro co n ro ro ro si -J- oi σi cn cn cn 4- cn ro 0 0 cn
N> — co ro — 01 ro ro — ro cn
4- co ω o o cn ω o cn o si c» si c» cn si si si ro ro to ω o ιo o — ω ιo ^ ^ 4- 4- 4- 4- 4- ω cn 4- 4- 4- 4- 4^ sl CD fO CD -' Cn 4- 4- C» -k Cn ^ O CD C» — sl CO Cn sl sl — ^ CD f04- C04- -i C CD 01 sl CO CO CD sl W 34- C — co oo ro co ro co CD CD CD CD CD 4- W 4_ — CD OI CD CD 4- — cn 4- — cD io si ω cπ ω cD co 4- ω cn ro » ∞ 4- ιNi cD cD si w co oo oo σi co oo ro si -' ro ω
Table 4
997089.7.dec 1575796H1 1222 1399 8 997089.7.dec 899613H1 1483 1735
997089.7.dec 3109970H1 1224 1502 8 997089.7.dec 755318R1 1161 1689
997089.7.dec 5732514H1 1223 1506 8 997089.7.dec g825993 1172 1536
997089.7.dec 927100H1 1232 1493 8 997089.7.dec 3821487H1 1174 1287
997089.7.dec 1731008H1 1504 1734 8 997089.7.dec 3293579H1 1182 1416
997089.7.dec g1740574 1510 1685 8 997089.7.dec 3441662H1 1183 1410
997089.7.dec 5527051 H1 1524 1766 8 997089.7.dec 1239073H1 1191 1432
997089.7.dec 6398114H1 1523 1751 8 997089.7.dec g3840776 1413 1762
997089.7.dec 4126433H1 1524 1793 8 997089.7.dec 4871671 H1 1433 1701
997089.7.dec 5186537H1 1529 1706 8 997089.7.dec 2253321 H1 1808 1876
997089.7.dec 5432074H1 1532 1779 8 997089.7.dec 1794040H1 1808 1876
997089.7.dec 3942839H1 1537 1810 8 997089.7.dec g2018816 1823 1876
997089.7.dec 4439832H1 812 1020 8 997089.7.dec g685621 717 1020
997089.7.dec 3570124H1 813 1111 8 997089.7.dec 2700859H1 723 995
997089.7.dec 3330642H1 813 1081 8 997089.7.dec 3513894H1 749 990
997089.7.dec 4332312H1 576 838 8 997089.7.dec 292713H1 1293 1557
997089.7.dec 6521581 H1 616 988 8 997089.7.dec 1864784H1 1302 1566
997089.7.dec 5656215H1 639 901 8 997089.7.dec g942975 1310 1616
997089.7.dec 3336136H1 652 895 8 997089.7.dec 2360739H1 1326 1567
997089.7.dec 1428142F6 669 1144 8 997089.7.dec 5731187H1 1335 1589
997089.7.dec 1428142H1 669 919 8 997089.7.dec 1733703H1 1336 1552
997089.7.dec g389543 708 1100 8 997089.7.dec 5690180H1 1339 1529
997089.7.dec 3668538H1 713 997 8 997089.7.dec 1863109H1 1702 1850
997089.7.dec 5042704H1 714 949 8 997089.7.dec 2102105H1 1694 1850
997089.7.dec 3699552H1 717 999 8 997089.7.dec 4339688H1 1703 1850
997089.7.dec 5156683H1 51 295 8 997089.7.dec 3817576H1 1694 1850
997089.7.dec g2020020 84 533 8 997089.7.dec 5353591 H1 1736 1850
997089.7.dec 6302085H1 228 530 8 997089.7.dec 3798035H1 1736 1850
997089.7.dec 4444078H1 288 515 8 997089.7.dec 4399365H1 1737 1850
997089.7.dec 5623338H1 507 827 8 997089.7.dec 4521845H1 1737 1850
997089.7.dec 4329118H1 576 821 8 997089.7.dec 1629818H1 1748 1850
997089.7.dec g698738 1498 1841 8 997089.7.dec 6480664H1 1678 1879
997089.7.dec g698716 1499 1824 8 997089.7.dec 3055643H1 1674 1850
997089.7.dec 3796755H1 1391 1705 8 997089.7.dec 707490H1 1680 1849
997089.7.dec 3772065H1 1393 1683 8 997089.7.dec 1945259H1 1682 1863
997089.7.dec 2921912H1 1400 1679 8 997089.7.dec 3127913H1 1682 1850
997089.7.dec 1338386H1 1406 1640 8 997089.7.dec 1483348H1 1682 1850
997089.7.dec 2208068H1 1792 1876 8 997089.7.dec 554627H1 1686 1850
997089.7.dec 2909601 H1 1797 1876 8 997089.7.dec 704029H1 1686 1782
997089.7.dec 2415770H1 1800 1855 8 997089.7.dec 1004810H1 1625 1834
997089.7.dec 3857026H1 1801 1868 8 997089.7.dec g616388 1637 1850
997089.7.dec 2455789H1 5 224 8 997089.7.dec 982190H1 1652 1850
997089.7.dec 6602471 H1 5 143 8 997089.7.dec 3803295H1 1652 1782
997089.7.dec 924429H1 1 170 8 997089.7.dec 3750923H1 1656 1850
997089.7.dec 6477763H1 1 564 8 997089.7.dec 1533117H1 1586 1800
997089.7.dec 3687663H1 9 307 8 997089.7.dec 5042002H1 1597 1845
997089.7.dec 4423309H1 22 307 8 997089.7.dec 2103420H1 1601 1729
997089.7.dec 5044370H1 22 287 8 997089.7.dec 1628004H1 1585 1791
997089.7.dec 3834620H1 23 302 8 997089.7.dec 761529H1 1610 1849
997089.7.dec 659525H1 1363 1640 8 997089.7.dec 4212159H1 1612 1859
997089.7.dec 5527891 H1 1364 1473 8 997089.7.dec 3846656H1 1623 1850
997089.7.dec 2532712H1 1377 1691 8 997089.7.dec 5546414H1 1194 1388
997089.7.dec 4321548H1 1379 1653 8 997089.7.dec 4073632H1 1202 1491
997089.7.dec 1427885H1 1437 1675 8 997089.7.dec 1703681 H1 1209 1394
997089.7.dec 4063504H1 1443 1706 8 997089.7.dec 4441637H1 1212 1407
997089.7.dec 2326755H1 1447 1685 8 997089.7.dec g1988696 1214 1646
997089.7.dec 1793547H1 1454 1753 8 997089.7.dec 5677789H1 1216 1461
997089.7.dec 755323R1 1462 1850 8 997089.7.dec 2019819H1 1216 1396
997089.7.dec 755323H1 1462 1675 8 997089.7.dec 4515993H1 1341 1551
997089.7.dec 3942852H1 1538 1809 8 997089.7.dec 4420186H1 1348 1596
997089.7.dec 3943170H1 1538 1803 8 997089.7.dec 1995894H1 1349 1588
997089.7.dec 4823018H1 1541 1686 8 997089.7.dec 5338351 H1 1562 1810
997089.7.dec 6166630H1 1562 1850 8 997089.7.dec g1716174 1562 1721
997089.7.dec 4049271 H1 1351 1534 8 997089.7.dec 6166622H1 1562 1876
997089.7.dec 3674696H1 1357 1479 8 997089.7.dec 1483801 H1 1562 1842
997089.7.dec 4144117H1 1464 1752 8 997089.7.dec 1294123H1 1567 1796
997089.7.dec 5733971 H1 1470 1729 8 997089.7.dec 4770575H1 1570 1845 CD CO 'v- OO O CD OO S CO — LO CO CD ^ CO CO O S ^- CO CN S CM Ul O 't S CD CD CO O — S O — S S OI CO — CD CO CD — CD — O CD >- S LO O » eθ S O tI) CI) ι|-
CI) (D O N O ^ W LO - n >- W !D ro O ιn N ω (O O IΛ O - 00 _ ι- S ^ - ^ ffl β β ^ CD LO O ^ * _ _
OO f CM lj- CD CM '^- ^l- CD CD S CD OO CD LO CD CD OO S CD S S O S CN CD CO OO CO CO '^- CD O '— S OO CD S CD S CD S S S OO S S CD S CD OI CD CD CD CD CD CD CD CD O CD OO CD CD CD CD O -3- CO — — — — — — — — — — — — — — — — — CM CM f CM — — CM CN — — CN CM CM CN CN CN CN CN CN CN CN CN CN CN CM CN C CN CN CN CN CN CM CM IM CN CN CN CM CN CM CM CM CN C
_
JN T — — O S -v- CD CN ^- S S O O O O O ^r LO OO CO CO LO — CN CO -t CD CD CD S O CM S i- j- Tr O LO CO CO CO S — CO Ol O Ol O - S n 'J- ft lO CD Ol 'Tf CD CD O CD S S CD CD CD CD — CO ^t g- CO S O O CD LO CO CO CM CM CM CO CO ->3- '* Ul U) U) CO CD CO S S M- -^- -<t "* U) Ul Ul CD S OO Oθ αθ CD CD O O — C
© n co o o o o w i- n c ^ c n n t ^ ^ ^ ^ M- to io u. co co ffl ffl O lO eO O) ^ * ^ ^ ^ ^ ^ ^ ^ ^ ^ ^ ^, ^t ^ !0 (O IO !0 (0 - _ - _ - tO (D _ _ S S © CN LO CD — — — — — — — — — — — — — — — — — — — — — * Ul
H CM CO — "p- "" *— r~ CD — — — — U I I LL. I I I I I I I — ri ri l l l l l l l cD l — l LO Oi co I S j- I I I h- I I CN I I I I I I- CD I I I DC ^ cs r CD LO — t CO I — ui — r: X α. C CO CN CO CO '3- DC I I — CO LO ^ CD — f CO CD S S CD S S CN O CD LO t" LO CN S CM — CO CM CO CD S '3- OO CM CO 'ςf 'd- O S S l LO CO l S O CD CO l— lloi ^ CO — CD O LO OI CD OI O CM CO OO OO S CO ^I- CO CN CM — 100) 0000) 00 — CD CM j- CO — — S — CO UI CO CM O CO CO CO CD S CO -*!" CO S O CD LO 't S — — — 00 CO LO CO LO CD Ol — — — CD CM CO S CN CO O S O LO CN CN — LO CO — CD LO CD CD CO OO OO S — CO CD CD LO OO S CO LO ^- CD ''*- — O O) CO S O 't CO S S CD O 01 Ul CM C O — Ol O Ol t Ol Ol CD LO — CD O OO OI S ^- ^I- CO CO CO I- CO IO CM CD CJI O O CD S LO — CO sl- CD CN CD CO — — CD — CO S OI S CO CO O CO OO C CO O OO UI CD S LO LO ^ CD — — CO CM ^ ^- S — OI CD CN CM T O CO S CO S st S CD CD S LO — CD CD O CO O S CM OO 't ' t — CD CO CO CO 00 CO U OO O O CD CM OO S S CO — CM LO -— LO S ^- CM CD — — CN CN — LO CM CO — CO f OO CO CO CN — CD CO CD CD CO CN S CD O st OO CO — — * inmsu CO LO CO CN C LO OO OO — CO CN CD CO LO CN *!- CO— CM CN CO CO COCN CO COCM t CO — "* CO CDCO — CM CM CO CN CO f — CN — CO CO COOO o o o o o o o o o o o υ o o o o o o o o o o o o υ o φ CD CD CD φ φ φ φ CD α> CD CD φ φ φ φ φ φ φ φ φ φ (D CD CD φ o o o o o TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ J TJ TJ TJ TJ TJ TJ TJ TJ TJ J J TJ XJ XJ XJ q q q q q ssss ssss ssss s s s s s sssssssss cb cb cb cbcb c
Ul Ul LO LO LO LO U Ul LO LO IO LO IO LO LO LO LO LO LO LO LO LO LO l l 000000 CO 00 CO CO 00 00 CO 0000 000000 CO CO CO 00 CO CO 0000000000 CM CN CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM sssss
CM CM CM CM CN CM CM CN CM CO CO CO CO O CO CO O CO CO CO CO O CO CO CO CO CO CO CO CO CO CO CO CO CO 0 0 00 00 CM CM CM CM C CN CM CN CM CN CN CN CN CM CM CN CN CM CM CM CM CM CN CM CN CN o0 o0 o o o o o ssssssssssssssssssssssssssssssss
oooooooooooooooooooooooooo
_- ) ca E-i
CN S CO CO 'sf S S CO S S S S O — sf 'st S CO CM S LO LΩ CO S — S S S S CO U1 U1 U1 CD CD CD CO S 00 C00000 UI UI CD CO OO CD CO CO CO CO CO CO CO CO CO CO CM O CD O O — LO O CD CD S CD Ol O Ol i- CN OO — CN CD — CO CO — LO O O - U1 U1 U1 U1 LO S S S S S S S S S S S S OO CO OO Ol CO lO CD CD CD CD CD Ol CD CD Ol Ol S sr cO — CN CM CM OO O O O O O O — CM CD CO CD CD αO CO Ol Ol 'St CD CN S S CN CN S CD — — CO -st Ul LO LO Ul LO LO LO lO lO Ul Ul CD S S CO OO OO CO CO — — — — — — — — CO CD CD CO CO CO CO CO 'St 'St in LO LO CD CO OO
— CD — — — — — — O I DC I I I CO —
I I I Li. I I I I - I II CM CO JO ΣZ CM CO S CD CD S 00 CM — -ϊϊϊlϊlDlOlϊϊ stir:
— LO CO f 00 CN CN CD Ol — 01 CD CD I st st O st O X I — I— sf CO CO S st CM S LO CO — CD W S CO st CO Ol I CDcocostcNstcooos — co " " — OO CM S — 00 00 CO o o CD — I
— "cf 00 CO Ul CM CD CO S CO S CO 00 CO O rl- CO CO 01 CM CM CO CN O CN LO — s co u — σi σi σi — — — ui — ui oo oo s s u oo CO Ol CO CO O O CO CO S 00 O CO CD CO CD CD o — ssso — oooosoσiσi
CO Ol Ol 00 S LO OO 00 CO 00 S — — O O — CN ^r S LO OO CN OO S CN CN — o CM S st st LO CD CD — CO CO CM CO LD — CO CD CD st CD sf st o c t LO Ul CO LO O 1- l" o U O CO — CM it 'l- Ol 00 U — — CD — S CO — n oi o ^ co CD S -o ω o t ^KD CM CM CD CD CO CD LO CO CO — S CD CD CO CO O CO CD Sf Sf 00t O COo s i CO 00 CO CO CD CN 00 st LO ioo S 0000 CD CD CO CO CO CD CO CO CD — o co co cN s st cN CN CN co st co — — st CO O CO S CO OO CO CD CD S S S CO st CO st — COCO O CD — C "tf CD Ol 00 CM CO st CD st LO St St CO LO CO CO CM CO CO — — Ul — CO CO LO st st st CO CO — — — — — — — CO CO D CO OO LO CO CO st i- CO CO CD CO CO CD CM CD — S CN CO LO LO LO CM LO CO C CD — — CM CO LO LO st — LO — — CO CO — 1- CD CD COCO CO COCO COCO CO CO CO CO CO— sf st — CO CO COCO CDsf sf — — COCM sf s o o o o o o o o o o o o o o o o o o o o o o o υ υ o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ o _ TJ TJ TJ TJ XJ TJ TI TJ TJ TJ TJ TI TJ TJ TJ XI XI TJ TJ TJ TI T) sssssssssssssssss — — — —' — —' — --—' — — — — — — — — —'. — -- ,- — --^T^T-:^ — T^ — — — sssssssss ss s's'ss's
S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S CM CM CM CM CM CM CM CM CM CM CM CN CM CN CM CN C o CD OI OI OI CD CD CD OI OI CD CD CD CD CD OI OI OI CO CO CO CO CO CO CO CO CO CO CO C CO CO CO CO O CO CO CO CO CO CO CO CO CO M cD σ σi σi σi σi σi σi σi σ σi σ σi oi σi σ σi cM CN CN CM CM CN CM CM CN CN CM CM CM W CN CN CM CM CN CN CN CN CM CM CM CM CN W
OOOOOOOOOOOOOOOO
OO OO OO OO OO OO OO OO OO OO OO CO CO OO OO OO OO CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD OI O CD O O O OI CD CD CD
CO CO — CO CO O O CN — CO OO st CN CO S OO st O OO OO S sf st CN CN CD in cO OO S — — — CO OO LO CD st Ol CD S st CCMM CCOOLO — COO — s τ- 00 ι- sf — CO CΩ O C CO CO CN CN '>- CO CO S — CD O CD O CN CO sf sf CD S st — st Ul CO S CN CD CO CO Ol st oO OO LO st CN LO Ul CD st CM OO st Ul S OO OO O CD CN U "-l C~N S CO CM O CD C S — CO CO O LO CO sf CN — CM LO CD — LO sf CM CO CN CD CN st cO CO CO CO CM sf CO CO CO CN CN CO CO CO CO CM CO st CO CO CO CO S CN CO CO sf CO CO CO CO CO Ul CN CO CO st sf st CO st sf st sf C CD C» S S CO S CO CO CO CO CN CN CN CO CN CN CN <N CM <M CN CM CN CN CN CM CM CM CM <N ( CM CN CN CN CN CN CN W CM CM CN CM CM C
_
OO OO OO CN CO st CO CD OO CM st CO CO LO LO — — CO S CN UI CD OO CD CD CD S CD OI OI OI — CO sf CO CO S S O S S CO OO CO LO r- O t- CN O O CM st O sf ol Ol Ol Ol — CN CN CM CM CM CN CM CN CO CO C CO CO CO CO st st lO CO CD S S OO OO CD CD C
© st sf st st s S CO OO OO CO — O O O O O O O O O O O O O O O O O O O O — — — — — — — — — — — — T- — 1- -- T- 1- -- — — — — — — — — — — — © st Ul Ul Ul Ul Ul lΛ UI UI UI CO — sf S — CD — CN CN CN CN CN CM CN CN IN IN CM CN CN CN CN CN CM CN CN CN CM CM CN CN CN CN CN CM CM Ul
H — — CM — — CD *- U X X X X I I I I 11 τz ii CO xCM uCO. ICO LO CM CD CD I I X X
CD LO S S — — CC CD CD CN CD ι cDιcoiuic —Mi —ι COιCD∑I:iCDιOIιCN I_ziCDιOιUIιOιCOι-l stl — Xr:icorI:rI:ico-ω-rl: i CDi α. CD CD CM CD _ I- x CN OliCOcCDrlrIicOiCOiT-rIicOiCOistiσiiCMoL CO — — CO CO CD CM sf sf 00 sf CM CD LO CO — sf — sf 0000 LO CO S O O sf CM OO O CO Ul CD sf S CD S CO O Ul Ul CO OO CO CM CO CM O sf S CO CO — sf st LO sf CO CD CO OO S LO CD S S O O S 00 S — CO CM CD Ol 00 CO O sf st 00 CO Ul U O Sf CM f — S CD S — S S sf CD OO CO S CD - CO sf CM LO O CM S CM O CO O CN CM st s CO CΩ S S O O CO CO CO CO CO CD OO OO O C
— sf S CD CO CO CD S S LO CD — CN O sf CO — O CO sf S S 00 CM CN CO O CO st CN — S CD CO — CM — CO LO sf S OO CD — CD O CD O CO CD — sf sf — — CO — S CO sf sf OO S CM OO CO C 00 — CD Ul — LO CO CO O CO sf CM CD CO CD O Ul 00 O CO sf st Ul CN S CD CO — — — CO O S CD sf O CD CO Ol CM OO CD - O O Ol O st cO S CO O O — — sf CD CO st 00 01 — O CO O sf 00 CO — l CD — CO CO CO CO LO sf sf U) sf — CM CO CO CO CO 0000 CD — O CO r- LO OO CD — CO S O CM S O — OO OO Ul sf S Ul OO Ul CO OO CD CM CO CO CN CN st LO S CN S OO CO S — CO CO C sf LO LO CO CO LO — — — — CM CO CO CO LO CN sf — CM 00 CO cost Ul CO sf CO COst 1- CO CO st CM Ul CO st LO st CM L LO — — 00 LΩ 00 CD LO — CD sf sf - — Ul Ol CN st — CO LO CN st — CO o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o s CNsCN CsNsCNsCNsCN CsM CsN CsN CsNsCN CsN CsNsCN CsNsCN CsNsCNsCNsCNsCN CsNsCNsCNsCNsCN CsNsCN CsNsCMsCMsCN CsN CsMsCNsCN CsN CsN CsN CsNsCNsCM Wssssssssssssssssssssssss
CsN CsN CsN CsN CsN CsN CsN CsM CsN CsN CsN CsN CsM CsM CsN CsM CsN CsN CsN CsN CsN CsN CsN CsN CsN CsN CsN CsN CsN CsM CsN CsN CsN CsN CsN CsM Wssssssssssssssssssssssssssssss
CM CN CM CM CM CM CN CM CN CM CM CM CM CM CM CM CM CM CM CM CM CN CM CM CM CM CM CM CM CM CN CN CM CN CN CM CM CM CN CN CN CM CN CM CM CM CN CN CN CN CN CN CN CN CN CN CN I CN CN CN CN CN CN CN CN CN CN CN CM CN CN CN CN CN CN CN W f
H
I CM — — — — — Ul — CD — O CO CD st CN CD Ol st Ul S — st CM CD S sf — CO Ol CD O CD CO CD S OO st
— T- Ol CO CO sf ssff sstf ssff ssff ssff sstt UUll ssff sstt sstt ccOO OO sstt sstt -- ssff ssff COOO UUll OO -- OO SS —— CCOO UUll —— ssff UUll OO SS OOOO SS SS UUll OO OO LO CO CD CD sf O CD LO CO LO CN CM - CD CO S O CD CD CD O S — — C OO Ol S CD CD CD l CCDD CCDD CCDD CCDD CCDD CCDD CCDD CCDD CCDD CCDD CCDD OO CCDD CCDD CCDD OOll OOll CCMM CCMM CCMM CCOO CCOO CCMM CCOO CCMM CCMM CCOO CCOO CCOO CCOO ssff CCMM CCMM CCNN CCOO CCOO CCOO CD S CD CD CO sf S S CD O O O CM CD - CO S CD sf — S CO CO i C CMM CM CM CM CM CN CM CM CM CM CN CN CM CM CM CN CM CO CM CM CM CM CM CN C IM CM CM CN CM M CM CM CN CN CM CN CN CN CN CM CN CM CM CM CM CO CO CO CM CN LO CN CO st cO LO LO CO 0
— CM CN CM CO LO CD S OO S LO OO UI CO CO S — — CD Ul — — — CO Ul Ul lO S CN CM st CO — CO CD OO OO CO OO CM UI S OO (M CN CN CO CO CO CO CO CO LO CD CD S CO OO OO O O O LO CD CD CO CN CM CN CN CM CO CO CO CO st st st st st st sf Ul Ul Ul Ul CN sf O CD CD S CD S
S S S S S S S S S S S S S S S S OO OO CO CO OO OO OO O O O O O O O O O O O O O O O O O O O O — CO S CD CD C CM CM CM CN st cO CD CO OO st st S S CM — CM CM C CM CM CN CM CM CM CM CM CM CM CM CN CM CM CN CM CM CN CN CM CN CM CN CN CM CN CM CM CM CM CN CN CM CM CM CM CM CM CM CM CM
IlL
o o o o c> o υ o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o co co co co co co co co s s s s s s s s s s s s s s s s s s s s s s s s s s s s s s s s s s s s s s s s s s s s
O
OO O S CD OO st OO OO — sf oO S CO sf S Ol CD CD — CO CN — CM CD st CO O — — CM S — CO UI UI CD S sf LO S S — LO CN sf CD CO S C
CO O — LΩ LO sf sf CD LO LO CD — LO sf CD sf LO Ol Ol LO CM CM st — sf S CD S S sf - CO sf CN S CD st st CN S S LO Ul CD O CO CO OO CM CO CM O O CD CO st - CD OO O LO CD sf OO S LO C
CO O) 01 CO CD CO S OO CD I_ CD CΛ IΛ CD CD CD CD LO LO Ul CD CD CD CD CD CD CD CO CO CO CD CD LO CO CO cn CO CD CD CD CD CD LO lO CD C s^ i- T- — — — — — — — — — — — — — — — (M CN CN CN CN CN CM CN CN CN CN CN CM CM CN CM CM CN CN IM CN CM CM CM CM S OO CO OO CD OI OI O OO — — — CD OO OO i- — — — — — — — —
_
CO CO sf Ul O O CO st OO CO CO CM CD O sf CO CD — sf LO sf OO CM OO O O O O O O sf OO - OO OI CN OO CD CO CD CD CD
CD CO CO CO st st sf st Ul CD CO S S CD CD CD CD sf sf st LO LO CD CD S S S S S S S S OO CO CO CD CD CD O - — — — st sf CO LO LO CD - S — — — CO O O — CM — O CN OO sf CΩ CD
© LO CD CO CD CD CD CD CD CO CD CD CD CD CO CO CO CD CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO st st st st CO CO sf lO S ffl rø © — — — — — — — — — — — — — — — — — CN CN CN CM CN CN CN CN CN CN <M CM <N CM <N CM CN CN CN CN CN CN CN CN CN CO CO CD CD CO CD CO S S S S S C» CO C» C» CO CO CO » Ul
H CD I- '"_ T- U I I I I l_ I I I X I I I I XIII I I — CO co r: 11 III I I I I 00 I t- I — X I I - _- I I i-Z- X I S CO 01 CD CO LO S sf CO O LO Ul S T- U) — O Ol H- CO I CM CO CM CM S CM CN CN S — CM O CO O LL IsIIsfICOI OII CDI I α. CO CM CNI UlI -—Xsf COIOI CMI CDI COI CM Ul 00 — O I CM CM o I 00 CO
CO CO — O — — sf O CO CM LO 01 LO LO CO CD CM 00 O O CO CD CN LO 1- S st CO CD 01 CO 00 00 S CM S CD co co l CO sf LO CM st CO CO CO CM CM CM —* CN 00 U st CO sf Ul CM S — st C sf CD 00 CD CO CO LO s CD LO S CN sf — O sf S f sf CD LO CO CM CO O LO 00 O LO O 00 — CO CD CO sf O O CO s 0
CD sf LO CN LO 0 0 CD st CM U) CM O 00 — O — CM st CO O sf
— CO sf LO CD Ol LO CO CO CD CO CD O CO CO 00 CO — S CM CO 01 CO CN S sf CD st s 0
CO S 00 CM LO S O 01 CO st CD 0 CM — CO
— sf LO 010000 CM CO CD st CD CD CD CO S 00 CD — sf CO O S O 0 Cs CO CD CD CD CM LO st S O st Ul CD CD CD S CO CO 00 0001 CO 00 o CO CD CD O S O 0 -— 00 O sf CD S CN CO s 0
O LO 01 01 LO sf 0 S S CO OO C Ul sf Ul S CO CD sf CN S O LO Ul — CM O OO S CM LO — CD — 00 LO s s s s S S sf CD sf 1- LO CD CO 0 s CD st S S CM CO S CD U sf CM *— sf O CO CD Ol CD CO O CO 01
U CN ss O S S — CD CO sf — Ul CO C CO CO CO sf — — Ul CO CO CO CO CN CO Ul CN sf CM sf sf 0)CM COCO CO LO U Ul LO Ul Ul CO CN LO OICO — CD CM — LO CM sf st CM st — — CO sf sf CO CN CO t Ul st st CDCD CO Ul OICO CN CO C o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o υ o o o o o o o o o o o o o o o o o o o o o o q o q q q q q q q q q q q o o o o o o o o o q q o q q q q q q q q q q q q q q q q q q q q q q q q q q q q q q q q q q q q q q q
-.- -(θ(ό - -ffltdd_tD_ιoιoιd_ - _ _ _ _
CrsNsCN CsM CsMsCM CsMsCMsCMsCM CsNsCN CsMsCMsCMsCM CsMsCMsCMsCMsCMsCMsCM CsMsCNsCM CsMsCM CsMsCMsCMsCN CsMsCM CsNsCMsCM CsMsCM CsMsCMsCMsssssssssssssssssssssssss
CsM CsM CsM CsM CsM CsM CsM CsM CsN CsM CsM CsM CsCsM CsM CsM CsM CsM CsN CsN CsM CsN CsN CsN CsN CsN CsN CsM CsM CsM CsM CsN CsN CsN CsM CsM CsM CsM CsMsssssssssssssssssssssssssss
CM CM CM CM CN C CM C CM CM CM CN CM CM CM CM CM CN CM CN CN CM CN CN CN CN CN CN CN CN CN CM CN CM CM CM CM CM
CN CN CM CM CM CM CM CN CN CN CM CN CN CM CM CN CN CN CM CM CM CM CN CN CN CN CN CN CN CN CN CM CM CN CN CM CM CM t
— 3
CO S CO S S CD CD CD S Ol CD CN CO sf cO CO LO CO O st Ol i- CO CD Ol st CN CD CD O OO CD st st O S O lO O CN lO sf O sf CD — CΩ CO CO τ- CO CO CO st cO CO O — CM — CM 00 CO O CD — OO CN CM sf OO S S CD CD CD CD CD S CD LO S - CD CM CD OO O CD — CO S S CD S S st O LO CN — — st CD LO CD CD Ol S LO CD T- OO st Ul Ol S S CO CD st CN CD CM Ul CO CN CD CD — sf CD L st sf sf st cO CM CO st st CM CM CM CN CM LO — CN CN CN C CN CM CM CO CO — O O O O CN τ- τ- — τ- τ- CM CM — — — — O CM — CN — CM CM CM CM O OO OO OO OO O CN CN CM CN CN CN CN CN IM CN CM CM CM CN CN CN CN CM CM CN CN CN CN N IN CN CN CN CM CM Ovl W
M Λ CO S OO OO CD CD S CO CD CO CO sf st cO CO CO CO CO CO CO CO CD O CN CN CN CO C CO st CO CO CO CO sf O sf sf CN CN st Ul Ul LO S O O O s^ CD Ol Oi σi σ CD CD O CN OO CO CD O O O O - — — — — — — — CN O O O O O O O O - ι- -ι- .|- CM CM CM CO CO CO CO CO CO sf st sf st U) Ul CD S S S S S CO CO S S S CO OO C
— — — — — — — CM CM CD CD CD O O O O O O O O O O O O O CD CD CD CΛ CD CD CD CD CD CD CD CD Ol Ol Ol Oi σi Oi σi CD CD Ol Ol CD CD CD CD CD Oi σi σi σi σi σi Oi σi LO Ul Ul Ul Ul U
CM CN CN CN CN CM CN CN CM — — — CN CM CN CN CM CM CN CN CM CM CM CM CN — — — — — — — — — — — — — — — — — — — 1- 1- — — — — — — — — — — — — — — — — — — —
i _- -: r: i i r:i i liii r: C cM i 1111 r: r: I - I x 111 I I I IIIOlIII I I I I r 100 r: 111
CM l I I Ul CO l CN CM CD Ol CO sf l DC s CD S S O sf X DC s UL Ul Ol Ul — s σi co s cN — — — sf CO O CO LO S LO S CO X: IO -S -sf 1st 1sf I CD I I O OO CO CD X X st X CM S I CO Ul S C S CO st CO CO CM CO — O OO CN CD — LO LO LO CN O IO CM CM CO CO — CD — OO st CD CO CO OO S O CM — O CD S S — f — U f OO sf O) LO S O CD — CD 00 — CD CO sf U CO CD O CD O — CN LO S CO CO CN O LO CN O CD CO CO O O CO S — S — — CN CM — CD CM CM O S st O CM Ol CO sf CO S S CM CM CD CO O O 00 o CO CO sf CM Ul CM OO LO O O CO CO sf CM CM O — CM CO O — 00 CD 00 O C0 O O — CO CM S — O O — 00 LO OO OO CM OO OO st st CO CD Ul OO st sf CM OO sf — — O LO O O sf LO CM CO — CD CD CO CO CD LO O CM sf CD U st CN OO O CO S O O 00 l Ul st Ol CD O CM CD CO CN LO LO O O — T- CN CO S CO CO LO S S CO CD st st lO O OO CO CN O CO CO CO CD CD S S l Ul st st st CM S — CO S LO 00 f S CO O CD 00 CD CD O CO CD CD CO S C CO st LO O CO — O CD — OI O LO CO CD CD CN CO CO CO — — LO LO IO O CD CO OO — LO S 00 U1 U1 LO L L — co co σi σi S CD CO CO LO CM — — CM — sf — O CO sf st S S CM — O CO O CD C CO LO CO OO — sf sf sf st Ul LO st st CD Ol CD — sf — sf lO — — CO CM sf sf sf CM CM C CN sf — — — CO CDsf sf CO CO sf CM CO CO CO sf CD CD CM M sf O)L0 sf CD CD — — CM CDCD LO U CO U
_ o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o υ o o o o o o o o o o o o o o o o o o o o o o o o o o o q q q q q q q q q q q q q q q q q q q q q q q q q q q q q q q q q q q q q q q q q q q q q q q q q q q q q q q q q q q q q q q q
00 _ d . - (D - - - . - ιo - ό - . C) - . .. co - _ _ - tό . _ - - _
CsM CsN CsNsN (sM CsN CsN CsM CsM CsM CsM CsN CsN CsM Csvl CsM CsM CsM CsNsN CsN CsNsNsN Wssssssssssssssssssssssssssssssssssssssssss
© s CMsCM C CN CsMs sM CsMsIM CsM CsM CsN CsN CsN CsM CsN CsM CsM CsMsCN CsN tsM CsN CsN CsN CsM CsM CsM CsM CsN C sM W o sNs CM CsM CsN CsM CsN C sM CsN CsN C ssssssssssssssssssssssssssss
CM CM CN CM CM CM CM (M CM CN <M CM CN CM CN CM CN CM CM CN CM CM CN CM <M CM CM <M CN CN CN CN CN CN CM CM CN CN C^ CN CN CN CN CN CN CN CN CN CM CN CN CM CN CN CN CN CN CN CN CM CN CN CM CN CN CN CN CN CN CN CM CN C
CD O — sf CD CD OO LO S CD Ul - — sf O — st C CD S CN O CO O Ol CD CM OO st st CO O — CM CN st CD CD CN CD O st S st CM CM CD OO CO OO — CD OO CD CD O CO S sf lO — CM OO CM CM CD CN IO O OI CD CD CD O O O O O O O CD CD O O O O O O O — — CM O O — — O — — O — — CM — — — — (M CN CN CN CN CM CM — — CN CM CN CM CN CM CM CM CN CN CN CM CM CM CN CM CM CM CM CM st st CM CO S st CD S CO - — — LO LΩ LO CM S LO OO CD S CD — — r- S S O LO CO OI O CD CD i- - CO CO CO LO Ul CO OO CD Ol Ol Ol Ol CD O O O CN CO CO CO st sf sf st st LO LO CD CD S S S S S S S S S S S S S S S S S S ∞ — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — 1- -- --
x x l OlII I II IIIII I I II IlLL lI I I IIIII I CM xCO x— sf CD 00 O O — O O CΛ Ol OO OO CM sf st st — CO S sf cO — S CD CM O Ol — — st CO CO CD 00 CD CD OI OO IO S S CO S - CD S OO sf — CD OO LO O — sf sf CM C0 C Ol -- Ol sf C0 Ul CO CN CD CO O — CO CM LO CD CO CN LO S CM CD — CD OO — CO CD CN — — O CD CO LO OO CO — S CO LO sf U CD CD CM LO CD CO CN — — sf CO LO — OO CD — CM S CD Ol CD Ol S Ol Ol CN OO CD CO CN i- CD St CD LO CD LO CO CN sf LO CN IO CD OO O LO lO S CO st — LO — CN CN O CO S CD — LO CD LO S O CM CO CD CM — CD CD CO S st S C CN LO CO CO S LO S CN — CD CD CD CD — CO LO CD CN LO CD CO CD CM — CO st CDUl CO CO — CM — — LO CM — LO LO LO sf LO LO LO — — LO LO LO LO CM CO — CM sf
CsM CM CM CN CN CM CM CN CM CN CM CM CN CM CN CM CN CN CN CN CN CN CN CM CM CM CM CN CN CN CN CN
CM CsNsCMsCMsCMsCMsCMs(MsCNsCNsCNsCNsCN CsM CsM CsMsCN CsMs(MsCN CsNsCMsCM CsM CsMsCMsCN CsNsCN CsM CsMsCN sCM CsNsC sssssssssssssssssssssssssss s CM CsMsCNsCN CsMsCMsCM CsMsCMsCMsCNsCM CsM CsM CsM CsNsCN CsM CssCM CsMsCMsCNsCMsCM CsMsCM CsM CsM CsM CsM CsM s CMsWssssssssssssssssssssssssssss
CN CN CM CN CM CM CN CN CN CN CM CM CN CM CN CN W CM CM CM CN CN CM CM CN CN CM CM CN C CN CN CN CM CM CM f
_ ! a H
O S O LO st cO CD S CO — CO — CD CD — — CO — ω — CD CO LO LO CD S CN CM CD CM CD — S CD O CD O O S UI O CD CN — CD CO S CD CD CN st τ- sf — S O — 'i- CD O CD O CD — — — O CO CO IΩ — LO O O vD O CO O LO CO CM CO st st cO lO CO CO st O st LO st st st Ol CO CO CD - sf CO CO O CO — lO sf CN CD st st CO st st lO O —
— CM CM CM sf C CN CM CN CO CM CN CO CO CO CO CN st CO S CN st st C0 st CD O100 CD CO O) CO CD CO CO O) Ol Ol CD Ol Ol Ol Ol Ol CO O) CO C0 CD CO CD Ol O100 CD CD CD 0O CO CD CD Ol Ol Ol Ol CJl CD
— — — — — — — — — — — — — — — — — — — — — — — — — CM CM CN CN CN IM CM CM CM Ovl CN I CM CN CM CM CM CN CM CM CM CM CM CN CM CM CN CN CN CN CM CN CN
CM CD CD LO S CO O S sf CD CM CO S S st st st CD CD CO CO OO CO CD CO CO CO CD O - LO CO CM CD LO st st CO CO CO OO CO CN CO CO CO CD — CD CD — — CN CO st CD LO OO OO LO O sf CO O C CO OO O O — — — CO sf st CO CO S CO OO CM CO CO CO sf st CD S S OO S OO OO OO OO CD CD CD CD O CD O O O O O O O O - — — — — CN CM CM CO CO CO CO CO CO CO CO CO sf CO CO OO
Ol Oi σi σ O O O O O O O O O O O O O - — — — — — — — U1 IΛ U1 U1 LO LO I— Ul Ul lΛ Ul CO lΩ ω CO CQ CO CD CO CO CO CO CO CO CO CO CO CO CD CΩ CD CD CD CD CD CO CO CD CO st st st
CD OI OI CD - — — — — — — — — — — — ^ — — — -- — — — — CN CM CM CN CM CN CN CM CM CM CN CM CN CM CN CN CM CM CN CN CM CM CM CM CM CN CN CM CM CM CM CM
I
— CO CM
— Ol CO Ul
_ 0 O υ 0 O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O υ 0 O O O O O O O O O O O O O O O O O O O O O O O o 0 O 0 0 O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O 0 0 O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O
90
CsMsCM CsMsCM CsM CsM CsM CsM CsM CsM CsM CsNsM CsM CsM CsN CsM CsM CsM CsN CsN CsN CsN CsN CsN CsN CsN C C C N CN CN CN CN CN I s N CsM CN CN CM CN CN C sN C M CsN MsM CMsM C CN CN C CNsM CNsCM CM CN sssssssssssssssssssssss
© CN sssssssssss o CsC M CsN CN CM CN CN CM CN CN CM CM CM CM CM CM IM CN C M CsN CsMsCMsCN CsN CMsCNsCM CsMsCN CsN CsNsCNsCMsCNsCN CN Csss N CN M CN CNsCNsCNsM C CN CsNsCNsCM CsNsCM CsM CsMsCMsCM CsMsCMsCM CsMsCN CsN Wsssssssssssssssssssssssssss CN CN CN CN CN CN CN CN C\1 CN CN CM CN CN CN CN < CN CN CN CN CN CN CN CN CN CM <N CM CN CM CN CM
Table 4
272721 .θ.oct 1399608H1 1878 2128 12 272721.6.oct 6429277H1 2266 2Θ55
21212' .β.oct 541556H1 1883 2120 12 272721.6.oct 6098047H1 2272 2445
21212' .6.oct g2017461 1891 2176 12 272721.θ.oct 2308651 H1 2272 2524
27272" .6.0C1 4727Θ56H1 1898 2168 12 272721.θ.oct 3014469H1 2276 2552
27272" .6.oc1 g3003890 2872 3204 12 272721.θ.oct 78e983R1 2277 2805
27272" .6.oct 564887H1 2875 3001 12 272721.β.oct 78Θ983H1 2277 2537
27272" .β.oct g2932091 2875 2996 12 272721.6.oct 830Θ95H1 2277 2537
27272' .6.oc1 g500166 2875 3008 12 272721.6.oct 1858762H1 2286 2549
27272' .6.oct 5525644H2 2908 3182 12 272721.β.oct g1026279 2293 25Θ4
27272' .6.oct 1881380T6 2928 3385 12 272721.θ.oct 4303854H1 2302 25Θ7
27272' .6.oc1 4941268H1 2228 2490 12 272721.6.oct 16Θ2959H1 2302 2534
27272" .6.oct 1730067H1 2238 2459 12 272721.θ.oct g1792740 2306 2458
27272" .6.oc1 1294012H1 2252 2504 12 272721.θ.oct 359Θ126H1 2312 2Θ01
27272" .6.oc1 1293975F1 2252 2827 12 272721.θ.oct 1689619T6 2307 2903
27272' .6.oc1 1293975H1 2252 2497 12 272721.β.oct 1820290H1 2317 2572
27272' .6.oc1 1272847H1 2252 2500 12 272721.e.oct 1820281 H1 2317 2573
27272" .θ.oct 2957128H1 2252 2545 12 272721.6.oc1 1873055H1 2324 2555
27272' .6.oc1 2439817H1 2255 2485 12 272721.6.oc1 3298163H1 2325 2582
27272' .6.oct 2040137H1 2255 2519 12 272721.θ.ocl 80574ΘT1 232Θ 2908
27272" .6.oc1 2264921 H1 2257 2464 12 272721.6.oc1 5527791 H1 2332 2605
27272 .θ.oct 2268728H1 2257 2487 12 272721.6.oct 3831404H1 2336 2550
27272" .6.0C1 1690407H1 2259 2432 12 272721.6.0C1 3840984H1 2338 2587
27272" .θ.oct 763923H1 2260 2536 12 272721.6.oct 805746H1 2338 2573
27272" .θ.oct 4894289H1 1176 1459 12 272721.θ.oct 2201471H1 2337 2578
27272" .θ.oct 2556183H1 1181 1433 13 461θ03.4.oct 26Θ7533H1 1512 1743
27272' .θ.oct 3319651 H1 1242 1519 13 461603.4.0C1 1862102H1 1523 1792
27272" .θ.oct 831550H1 1250 1452 13 461603.4.0C1 23554Θ5H1 1580 16Θ5
27272 .θ.oct 3118134H1 1275 1558 13 4Θ1603.4.OC1 g3θ45529 1645 2081
27272" .θ.oct 5902550H1 1283 1580 13 4Θ1603.4.OC1 3293621 H1 1670 1906
27272" .e.oct 4996858H1 1283 1563 13 4Θ1603.4.OC1 g2335995 1736 2089
27272 .e.oct 6113602H1 1295 1617 13 4Θ1603.4.OC1 3752905H1 1758 2019
27272 .e.oct 4518784H1 1297 1551 13 461603.4.OC1 6094776H1 1908 2151
27272 .θ.oct 1989226H1 1311 1571 13 461603.4.oct 418011H1 1922 2103
27272 .θ.oct 956338H1 1317 1413 13 4β1603.4.oct 417150H1 1922 2101
27272 .e.oct 46Θ9770H1 1322 1560 13 4ei603.4.oct 418870H1 1922 2118
27272 .e.oct 2821261 H1 1324 1630 13 4θ1603.4.oct 413059H1 1922 2129
27272 .β.oct 1567553H1 133Θ 1535 13 461603.4.oct 413059R1 1922 2399
27272 .e.oct 4880720H1 1352 1590 13 461603.4.oct 418870R6 1922 2275
27272 .θ.oct 3881948H1 1359 1622 13 461603.4.oct 417906H1 1922 2125
27272 .e.oct 315670ΘH1 13Θ2 1Θ47 13 461603.4.oct 696345H1 1964 2151
27272 .θ.oct 5040118H1 1380 1622 13 461603.4.oct 5694066H1 1987 2177
27272 .e.oct 3788560H1 1382 1675 13 461603.4.OC1 2637796H1 2019 2288
27272 .β.oct 3806Θ87H1 1389 1646 13 461603.4.oct 359Θ965H1 2058 2223
27272 .6.oct 1282584H1 1425 1685 13 461603.4.oct 2673521 H1 2077 2325
27272 .θ.oct 1703878H1 1441 1664 13 461θ03.4.oct 3405821 H1 2085 2349
27272 .6.0Ct 859573R1 1445 2007 13 461θ03.4.oct 482251 H1 2100 2335
27272 .θ.oci 859573H1 1445 1679 13 461Θ03.4.OC1 484926RΘ 2100 2452
27272 .6.oct 4848606H2 144Θ 1695 13 461θ03.4.oct 5693609H1 2114 2272
27272 t.θ.ocl 1558344H1 1446 1638 13 461603.4.OC1 g1188364 2133 2267
27272 1.θ.oci 3720883H1 1454 1628 13 461Θ03.4.OC1 3070461 H1 2171 2336
27272 i .e.oci 3720892H1 1454 1739 13 461603.4.OC1 g2779413 2210 2611
27272 1.θ.oci 026143H1 1456 1796 13 461603.4.OC1 3356374TΘ 2218 2437
27272 i .e.oci 5030608H1 1457 1723 13 4Θ1Θ03.4.OC1 g1780256 2225 2624
27272 1.θ.ocl 4674029H1 1470 1746 13 461603.4.OC1 1756956H1 2404 2651
27272 1.θ.oci 2634326H1 1484 1731 13 461603.4.OC1 1250711F1 2456 2798
27272 i .e.oci 5830504H1 1489 1704 13 461603.4.OC1 5005176H1 2456 2684
27272 1.6.0C1 5435568H1 1494 1715 13 461603.4.OC1 3001108F6 2554 3018
27272 i .e.oci 290356R6 1494 1974 13 461603.4.OC1 1252411H1 2570 2798
27272 1.θ.ocl 3356346H1 1507 1735 13 461603.4.OC1 1951691H1 2583 2801
27272 1.6.oc1 3342545H1 1531 1774 13 461Θ03.4.OC1 178Θ802H1 2625 2725
27272 1.θ.ocl 484755ΘH1 1552 1806 13 4Θ1603.4.OC1 1250711H1 26Θ8 2798
27272 1.6.oc1 g2156699 1574 1992 13 461603.4.OC1 3001108H1 2712 3018
27272 1.6.oc1 2503041 H1 1574 1797 13 461Θ03.4.OC1 3154469H1 2841 3111
27272 1.6.oc1 4883822H2 2265 2537 13 4Θ1Θ03.4.OC1 g4525001 1 365
27272 1.6.oc1 1542559H1 2265 2478 13 4Θ1603.4.OC1 g4148045 2 441
27272 l.θ.oci 6430113H1 2266 2655 13 4Θ1Θ03.4.OC1 g3846311 5 228
27272 1.θ.ocl 4341009H1 2265 2586 13 4Θ1603.4.OC1 g4 36972 5 3Θ5 CO CO CO - S S S S S sf O S C
, τ— r~
ui
_.
H
O sf CM OO st O — O CN CO — O CO CO CM S O — S LO sf Ul CM CD CD U s
CO st — O S CN st CO CO CM CD CO CO CN CO OO OO S — CD S — O st CN CO st O CO sf LO OO CD LO CD CD CO — CO OO CM CO CD O CD O CM O CD O CO CO CO LO O CO LΩ — — CD CO — — O sf sf CO O CD — CD CO CN O O CO LO CD OO st CM CD CD CD CN st — CM S OO O O O — OO O — O O — CO CO O S Ul lO O st lO CD O LO — 00 CO — CM Ol sf CO CD CO OO CO LO OO Ul st st CM - O CO st Ul LO CM Ul st CO Ul lO CD Ul CD S S LO CD LO S CD S S S S — τ- 00 — CD — — — — — — — — — — — — 00 — — — — — — — — CM CO — — CO — CN sf LO CN CD CD CD S — sf OO U1 00 LO — O S — CN
CD Ul LO O CM CD Ol OO CD LO O CO CM st CD CM O OO st st CO Ol S lO CD CN CO — CD OO S OO O O — O LO st CN OO CO LO OO CD — 00 CD Ol sf st O S Ol C C sf sf CD CO CM st σi — O CM CM CO sf sf LO st LO LO Ul S CD - OO CO CO CN LO CO O O — LO IO O — S CO CO CO OO OO CO st O sf LO st st S O sf O — S — sf sf oO CD CD CD S C CO st sf st CO - — — CN sf st st sf sf sf sf st sf st st st sf LO LO LO LO CD CD CD CD S Ol S σi — — S — — — S S — — 00 — CO τ- 1- OO Ol — — — CO CO S CD — — — CM CO sf S O
— — — —
— - LO X I LO I I I u_ — I I I I I I I — T T - I C I I sf I I ro T S T t T — T co o X X ii- X — X X I I I LL — .1
, — — o I X J. s s o X X s o o s — -
— s s o s o o o o ^ s ^ — - s o , — s
— s o s — s ^ s s o
Table 4
980541.1.dec g2525285 2467 2887 19 242082.10.dec g2359154 1208 1464
980541.1.dec g8831θ1 2494 2869 19 242082.10.dec g3179415 124Θ 1469
980541.1. ec g779712 2494 2615 19 242082.10.dec 673956H1 1283 1547
980541.1.dec g21θ3763 2508 2878 19 242082.10.dec 5482918H1 1286 1535
980541.1.dec g873216 2549 2733 19 242082.lO.dec 1395173T6 1287 1433
980541.1.dec g7θ4624 2544 2923 19 242082.lO.dec 2202416H1 1301 1558
980541.1.dec g876367 2544 2916 19 242082.10.dec 2202416F6 1301 1554
980541.1.dec g830021 2550 2897 19 242082.10.dec g1765329 1314 1441
980541.1.dec g756279 2551 2813 19 242082.10.dec g1190219 1368 1716
980541.1.dec g885095 2784 2898 19 242082.10.dec 600663F1 1369 1973
237996.1.dec g5056707 1 334 19 242082.10.dec g2589750 1404 1464
237996.1.dec 4531252H1 1 276 19 242082.10.dec 2133344H1 1408 1545
237996.1.dec 361421ΘH1 4 238 19 242082.10.dec 5024752H1 1428 1522
237996.1.dec 3111306H1 14 222 19 242082.10.dec 632098H1 1457 1724
237996.1.dec 3Θ04478H1 200 420 19 242082.10.dec 2202416T6 1466 1953
237996.1.dec 6152381 H1 355 630 19 242082.lO.dec 3091754H1 1480 1752
237996.1.dec 2937960H1 544 802 19 242082.10.dec 192Θ130H1 1487 1690
243267.9.dec 429897H1 1 168 19 242082.lO.dec 1926130R6 1487 1973
243267.9.dec 3277737H1 9 250 19 242082.10.dec 2292755H1 1505 1753
243267.9.dec 5052754H1 134 381 19 242082.lO.dec 2867724H1 1507 1805
243267.9.dec 59843Θ1 H1 203 400 19 242082.lO.dec g5438θ05 1519 1981
243267.9.dec 1754639T6 353 61θ 19 242082.lO.dec g728298 1521 1734
243267.9.dec 1754905F6 360 667 19 242082.10.dec g4003740 1527 1981
243267.9.dec 1754639H1 360 605 19 242082.10.dec g1θ78437 1539 1873
242082.10.dec 1962272H1 1 2Θ7 19 242082.10.dec g3445912 1548 1976
242082.10.dec 1710521H1 9 212 19 242082.10.dec g4188129 1548 1977
242082.10.dec 1710521 F6 9 361 19 242082.10.dec g4002844 1552 1974
242082.10.dec 924352H1 77 425 19 242082.10.dec g168θ101 1559 1976
242082.lO.dec 5121112H1 148 424 19 242082.10.dec 466459T6 1594 1932
242082.lO.dec 4270649H1 254 505 19 242082.10.dec 46Θ459H1 1602 1828
242082.lO.dec 1299353H1 286 498 19 242082.10.dec 46Θ459R6 1603 1973
242082.10.dec 3373518H1 288 559 19 242082.lO.dec 4Θ8718H1 1603 1828
242082.10.dec g2159501 304 511 19 242082.lO.dec g218θ213 1621 1988
242082.lO.dec 3748Θ35H1 314 569 19 242082.lO.dec 1220035RΘ 1671 1973
242082.10.dec g835896 371 613 19 242082.10.dec 38814Θ9H1 1682 1964
242082.10.dec g784665 372 46Θ 19 242082.10.dec g1266171 1694 1973
242082.10.dec 1374464H1 528 747 19 242082.10.dec 059532H1 1699 1905
242082.10.dec 5451856H1 569 799 19 242082.10.dec g1189494 1711 1974
242082.10.dec g1324135 592 1037 19 242082.10.dec g105θ493 1751 1988
242082.10.dec 5216316H1 659 906 19 242082.10.dec g3232610 1797 1985
242082.10.dec g1056590 757 998 19 242082.10.dec g122θ704 1839 1978
242082.lO.dec 6429773H1 777 1312 20 019239.1.dec g2037390 1 291
242082.lO.dec 2127293H1 844 1091 20 019239.1.dec 1297333F6 1 491
242082.lO.dec 5579027H2 868 1140 20 019239.1.dec 1297333H1 1 263
242082.lO.dec g1276071 900 12Θ3 20 019239.1.dec 3331976H1 407 597
242082.10.dec 3551092H1 905 120Θ 20 019239.1.dec 3040041 Fθ 523 770
242082.10.dec 1898869H1 918 1197 20 019239.1.dec 3040041 H1 523 788
242082.10.dec g5234067 993 1464 20 019239.1.dec 4938598H1 658 815
242082.10.dec g2159502 1005 1467 20 019239.1.dec g677095 684 892
242082.10.dec g5396677 1009 1465 20 019239.1.dec 503922H1 692 921
242082.lO.dec g5671070 1020 1464 20 019239.1.dec 503922R1 692 1098
242082.lO.dec g1384768 1038 1464 20 019239.1.dec 5318750H1 866 1056
242082.lO.dec g3785412 1049 1464 20 019239.1.dec 3Θ93495H1 888 1098
242082.lO.dec 600Θ63H1 1062 1346 20 019239.1.dec 3Θ93495FΘ 888 1331
242082.lO.dec Θ00663R6 1062 15Θ4 20 019239.1.dec Θ18665H1 1020 1287
242082.lO.dec Θ00663R1 1062 1545 20 019239.1.dec 2158372H1 1020 1273
242082.10.dec g2185852 1074 1330 20 019239.1.dec 1297470FΘ 1142 1583
242082.10.dec g1774698 1074 1455 20 019239.1.dec Θ89278H1 1142 1365
242082.10.dec g2141640 1120 1541 20 019239.1.dec 1297470H1 1142 1286
242082.10.dec g564444 1150 14Θ4 20 019239.1.dec 3775993H1 1142 1393
242082.10.dec 674332H1 1155 1408 20 019239.1.dec 1297470T6 1163 1789
242082.10.dec 4533148H1 1163 1414 20 019239.1.dec g2557145 1210 1404
242082.lO.dec 2867266H1 1164 1463 20 019239.1.dec Θ343003H1 1232 1451
242082.lO.dec 296929ΘH1 1166 1415 20 019239.1.dec 4769071 H1 1271 1509
242082.lO.dec g831436 1169 1428 20 019239.1.dec 3040041T6 1302 1754
242082.lO.dec 2754527H1 1202 1463 20 019239.1.dec 2840Θ27FΘ 1510 2050
242082.10.dec 3891815H1 1207 1439 20 019239.1.dec 2840Θ27H1 1510 1760 o o o o o o o o o o o o o o o o υ o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o υ o o o o o o o o o o o o o φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ xj xj xJ Xi xj j TJ TJ TJ TJ TJ TJ TJ T T Tj -q -q TJ TJ TJ XJ TJ T^ cb cbcb cb cb cb cb cb cb cb cb cb cb dcb cbcbcb cb d n n cbcb cbcόcb cb cb st sf st st st sf st st st st st st st st st st st st st st st sf sf sf st st sf st st sf st st sf st st st st st st st sf sf st lO LO LO LO LO LO lO LO U) lO Ul LO LO LO L LO LO lO _ CD CD CD CD CD OI CD CD OI CD CD CD CD ID OI O OI CD CD CD CD CD CD CD CD CD CD O O OI OI OI OI CD OI OI CD CD OI CD CD CD CD IO LO LO UI UI LO IO UI UI U^ cD CD CD CD Oi o σi σi σi cD cD σi σi cD CD CD CD CD CD CD CD CD CD CD CD CD σi σi σi oi σi σi σi σi σ σ σi σi σi σi cD CD CD co co co co co co co co co co co s s s s s s s s s s o o cD CD CD O oi σ oi σi oi σi σi σi σi oi oi σi σi oi o oi cD CD CD CD CD CD CD Oi oi cD CD CD Oi σi σ σ σi oi oi σi σi oi cD st sf sf st st sf st st st st st cD CD CD CD CD Φ
CO 0000 C0 00 C0 0000 CO CO <» 00 CO O0 00 CO O0 00 CO CO CO CO CO v» CO CO ∞ CO CO 0O CO C0 CO CO 00 OD 0D 00 000000 O0 » i- T- T- — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — T- T- — — — — CM CM CM CN CM CM CN CN CM CM CM CO CO CO CO CO CO CO CO CO CO sf st s
CM CM CN CM CN CN CM CN CM CM CM CN W CM CM CN CM CM CM CM CN CN CN CM CM CM CN CM CN CM CM CM CM CM CN CN CN CN CN W t ω
H sf CM — CO — OO OO OO OO CO CM S S CO CD CD CD CD CN O — CD OI CD S CO CD — — CO CD CO — UI O OI CN LO O CD LO OO OO CO UI O OO CD CD OI OO CD CD CD CO LO - CD sf O LO CD - CD C sf S — CN CN CN CM CO CM CM CO CD OO CD CN O O OO sf O — CM CN Ol sf Ol LO — S sf oO S O O CO S CM CO CD Ul - — OO LO CM O CD CD st O CO OO OO O CO sf S OO CM O S sf S OO C OO OO CD CN CO CO CO CO CM CO C sf S CM -1 — CN CN CO O CO CD S CO sf O CD CO O — OI CD CD O - O — CD CN CO st CN CO CD st S LO CO OO OO OO Ul Ul CM S CD OO CD CO S S Ul S S OO O CD — — Cvl CM CM CM CM CM CM CO CO CO CO CO CM st st S - — — — — — CM — — CM CM — — CN CN CN CN CN CN CN CN CM CN CO CN CN CM CN CM CM CN Cvl CN CM CO CN CM CN CM CN CN CN M CM CN CM CO CN C
S O OO — CO st CD CO O st O S O T- sf Ul st Ul UJ S S S CD - — CD CO CD CD CD O O — CN CM CM CD CN CD CM CN OO O S S S CO CO st st CD T- st oO OO CD Ol S st cO CO CO CO Ul sf CO sf Ul CO S sf O — CO — S S sf cO O CO Ul Ul lO CN CM CN CO S S — sf CN CD CD O CM CO CD CD Ol sf sf O CO CO CD CM CM CN CM OO CO O O O CO sf LO Ul Ul Ul Ul S OO CO U1 S 0O CD CD CD CD CD CD — CM CM CM CO CM CM CN CD CO CD CO CO CO S S S S S CD O O O O O S — CM CM CM CM CO CO CO CO CO S sf sf st st st sf sf sr st st st st st U — — — — — — — — — CO CO CO CO CO 1- — LO S — — — — — — — — — — — — CM — — — — CM CM CN CM CN CN CN CN CN CM CM CM CN CN CN CM CM CM CN CM CM CN CN CN CN C CM CN CM CM C T- ,_ — — — - — — — I X - oo I I I I co I I O O Ul CN S — — CN — LO CO CO OO — CO CO OO CD S S — — CD — CO S s CO O — — CO UI S CO UI CO — CN CM C CD LO sf OO — OO OO OO CD UI S LO CD C S Ol CD lO S CD CD CD st — sf lO LO OO OO — CD CM CM CM CM S S — LO LO C LO sf CD CD Ol LO 10 Ul sf CO COCO CO C
O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O o o o o o o o o o o o o o o o o o o o υ o o υ o o o o o o o o o o φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ o _ XJ XJ X1 XJ X1 XJ XJ XI XI X1 XI XI XI XJ XJ XI X1 XI TJ XI XI TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ XJ XI X1 XI XI XI XJ XI X1 XI XJ TJ TJ TJ TJ TJ T1 XJ XI X1 XJ XI XI XI XI XI XI XI XJ TJ TJ TJ TJ X so lD CD 01 01 01 01 Cn CD 01 rt CO (r) CO CO (r) CO C CO CO C CO C CO C CO C C C C CO CO CO CO CO CO CO CO CO O CO CO C^ CO CO CO CO crj CO σl CO C st st st st st st st st sf st st st st st st st sf sf sf st st st sf st sf st st st sf st st st st st st sf st sf st sf st st st st sf st st st s^ cN CN CM CM CN CN CM CM CN CD cD CD σi σi σ σi oi σi σi oi σi σi oi oi σi σi oi σi oi oi σ oi oo cD cD CD CD CD C o o cD CD CD CD σ σi σi oi σi oi σ cD CD CD c^
© σi σi o cD CD CD CD CD CD CD CD CD CD CD CD CD cD CD CD CD CD CD CD CD CD CD CD CD σi cD oi oi oi σi σi σi σi oi σi σi σi σi σi cD CD CD CD CD σi σi σi σi σi σi σi oi σi σi c^ o — - T- T- T- — - T- T- T— 0) 01 01 CD 0101 01 01 CD CD 01 01 01 CD CD 01 CD CD CD CD CD CD O) 01 01 01 0101 01 0101 01 01 01 01 0101 CD CD CD CD CD CD 0101 01 0101 01 CD CD CD 01 CD CD CD CD C o o o o o o o o o ω ra — ∞ co ∞ co — co co co ra co — oo co co co co ∞ ∞ ∞ ∞ ∞ rø
O O O O O O O O O — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — T- — — — — — — — — — — — — — — — — — — — — —
CM CN CN CN CN CM CM CM CM CM CN CM CM CN CN CN CN CN CM CN CM CM CM CN CN CN CN CM CM CM CM CN CM CM CN CN CN CN W
oφ oφoφoφ oφoφoφ oφoφoφ φooφ φooφoφoφoφoφ φooφoφoφoφ oφ oφoφ oφ oφoφ oφoφoφ oφ φooφ φooφoφ oφ oφoφ XJ sLO UslsUlsUlsUlsUlsLO LsOsLOsUlsUlsUl UslsUl UslsUlsUlsUlslO UslsUlsLOsUlslOsLOsLOsIΛ UslsLOsUlsUlsWsssssssss — — — — — — — — — — — — — — — — — — — — — — — — —— — — — — — — — — — — — — — — — — — OO OO OO CO CO OO OO CD OO OO OO OO OO CA CO CO CO CO CO OO OO OO CO OO CO CO CO OO OO CO OO OO OO CO ΑJ OO OO — 000000 00 sf st st st st sf st sf sf sf st sf st st st st sf sf st st st st st st st sf sf st st st st st st st st st sf st sf st st sf
OO OO OO OO OO CO CO CO CO CO CO CO CO CO CO CO OO OO OO OO OO OO CD CD OI OI OI OI OI CD CD CD CD CD CD OI OI OI OI OI O CD CD CD CD CD OI OI OI OI CD CD OT CN CN CN CN CN CN CN CN CN CN CN CN CN CN CN CN CM CM CN CN CM CN CN CN CN CN CN CN CN CN CN CN CM CN CN Cvl CN CM CN CN W
O S CD CO CD O CD S CD CD CO CO LO CO O S — CO LO O CD O CN CΩ IO CN — CD O S CO O CM CD CD sf CM 00 — S S 00 O — CD LO sf Ul — sf cO OO OO CN CN S st O -r CD — CO CM CM O sf S CN CO CD st S LO OO Ul Ul sf - CD st sf CD CO st O CM OO — OO OO LO sf Cvl sf CD CO CD CO O sf st CO CO CO — 0O U1 0O — CD sf S O CD LO CO CN O C 00 st CO st sf Ul st CO CN Ul sf Ul CN CN Ul CM Ul CO Ul CO st sf sf Ul CD sf CO CO CO LO UJ CO Cvl CM OO st CM st CO sf sf sf sf LO st st sf st sf Ul S CD CD CD CD CM CO CM CO C
CM
CD Ul Ul CD — CD LO O! CD CM CD st Ul CD CD CM S Ul O LO OO CD CO O CO CO CO CO LO CO S CO O O CN — CD LO CO CO O)
CN CO CO O LO O CM — CN CD CD LO UI CD CD S CD CN S — S CD O CO st O — CM CN CN CN CN OO O CD Ol S S LO CD CO CO OO st — O OO
— — — CN CO st — CN CM — — st CO CD CD - — — — — CN CO sf CO sf — — LO — — — — — CN CN CN CN CN CN CN CN CO CO CO LO LO CO CO CD LO OO CO — S S
CO CO CM »- - — — T- T- — co O — — OO — — — — CO CD T- - CD T- T- — CO '- CD '- CN - — - — — CO CO CM cO - CO O LL I X IL l I I I ID l β l l h lO I I DC Li. LL H I I I JO I U- I I I sf I sf I I X H S I LL I Jg - S li sf - I I I O LO CM I
CO CM S S st O S CN l— CO CO CO OO sf O 0 H0 — I s -f I — D OC I 00 I sf O CD OO OO st O OO CO CD st OO l LO CD CD CO LO CN OO CM OO — CM CM CM — LO CO CO X I O O CO I O sf C S CM LO CO CD C s S CO CD CN sf — CO O O S CD OO CD LO 00 S 00 sf CM sf CO UI CO OO OO — — OO S OO τ- 00 — — sf CM S CD CM CD CO Ul CD sf CO st O t sf s CN CN CN CΩ st oO — — S CM sf O CM sf O CD CO LO CO Ul O S st CO OO CD CD CO S OO CD S st oO OO — CM st sf Ul CO CO CD CD CD — sf Ul Ul S S O LO CM S Ul Ul CD CO — — CO U O O 00 sf 00 OO S CM O CD S sf — CO CO OO CN CO Ul CD LO CO sf st CM CN CO CD LO LO CD S O UI CD OO OO — — OO CD CO Ul S CD LO S S S sf S OO OO — CO S Ul sf — LO CO LO S CO — LO CM C o CD CD , — — CM OO — CD CD — CO CM CM CO LΩ CM CO CO 00 sf S 00 — CN CN S S — S CD CD S S CD CN CO CO CD CN Ul — CD CD CD O O O — CN CN CN sf CD Ol — sf sf — O O O CO CO CD S L O CO CO o O CD CO — CO CO CD CM OI OI S CM CD CO CO O CO CO S CD CO Ol CO CO S CO CD S CO CO sf — O O CD — LO CM CD — CM O O O CN CO Ul LO sf sf st S — sf S O CD sf t st CM — CM
CD CM st sf CM CO CM — l — CD— — sf CO — — COLO U sf — CM CO — — — Ol— COCO — — st CO sf sf LO st CD CDCM COCO CM CM CM CD CM sf st f sf COCM COsf CM sf sf — CO CO COCO CO o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o
_ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ o XJ XJ XJ XJ XI XJ XI XI XI XI XJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TI TJ XJ XJ so
r-' — — — — — — —' — — _ _ — _ _ — CO — > -> -> CO _) — — — — — — — cό _i
— — — — — — — — — CD CD CD CD CD Ol Oi σi Oi σi σi CD CD CD CD CD CD OO CO OO OO OO OO OO OO S S S S S S S S S S S S S S S S S S S S S S S S S st sf st st st sf st st
CD CD CD OI OI CD CD CD CD CN CN CN CN CN CN CN CN CM CN CN CN CN CN CN CN CN CN CN CN CN CM CM CN CM CΛ CD CD CD CD CD CD OT
© O O O O O O O O O CD CD CD CD CD Ol CD CD Ol Ol CD Ol Ol Ol Ol Ol Ol CM CM CM CM CM vM CM CM - — — — — — — — — — — — — — — — — — — — — — — — — CD CD CD CD OI OI CD OI
O O O O O O O O O Ol CD Ol Ol Ol Ol CD CD Ol Ol Ol Ol Ol Ol CD CD CD st sf st st st st st st O O O O O O O O O O O O O O O O O O O O O O O O O S S S S S S S S O CD CD CD CD Ol Ol CD CD CD CD CD CD CD CD CD CD CD Ol Ol Ol Ol Ol Ol Ol CD CD st sf st sf st st st st CD CD CD CD CD CD CD CD CD CD Ol CD CD CD C^ st sf sf st st st st sf st LO IO LO LO LO UJ Ul Ul LO Ul Ul Ul Ul Ul Ul Ul Ul CD CD CO CD CD CD CO CO S S S S S S S S S S S S S S S S S S S S S S S S S CO CO OO CO OO OO OO CO CN CM CM CN CN CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM CN CM CN CN CM CN CM CM CM CM CM CM CM CM CN CN CM CN
Table 4
481750.1.dec 3392402T6 1656 21Θ4 33 902791.3.dec 3429917H1 1821 204Θ
481750.1.dec 4181819T8 1701 2172 33 902791.3.dec 1449768F6 1855 21 Θ7
481750.1.dec g3960862 1725 2116 33 902791.3.dec 1449751 R1 18Θ5 21Θ7
481750.1.dec Θ051495J1 1763 2337 33 902791.3.dec 48586Θ0H1 1900 2159
481 50.1.dec 3514153H1 1774 2042 33 902791.3.dec 3519949H1 1920 21ΘΘ
481750.1.dec 1793831 R6 1795 2194 33 902791.3.dec 873395H1 1923 2104
481750.1.dec 1793831 H1 1795 2095 33 902791.3.dec g23507θ4 194Θ 21 Θ7
481750.1.dec 4129281T6 1886 2251 33 902791.3.dec 17Θ175H1 2052 21 Θ7
481750.1.dec 3725148H1 1901 2193 33 902791.3.dec 898974H1 1Θ01 1790
481750.1.dec 4820753T6 1929 2464 33 902791 ,3.dec g848299 1396 1723
481750.1.dec 1793831Tθ 1946 2458 33 902791.3.dec Θ512815H1 1040 1595
481750.1.dec g1745418 1953 2193 33 902791.3.dec 4374584H1 782 1043
481750.1.dec 6454265H1 1971 2424 33 902791.3.dec 898974T1 1921 2047
481750.1.dec g1241685 1995 2108 33 902791.3.dec g13θ770θ 1575 1842
481750.1.dec g4573702 2061 2494 33 902791.3.dec 141702H1 1563 1787
481750.1.dec 3705964H1 2069 2338 33 902791.3.dec 450088H1 764 995
481750.1.dec g21122θ8 2085 2495 33 902791.3.dec 3508596H1 1 212
481750.1.dec 3276661 H1 2251 2494 33 902791.3.dec 1390970H1 1 218
481750.1.dec 6051495H1 1163 1638 33 902791.3.dec 4575303H1 40 321
481750.1.dec 5814635H1 175 391 33 902791.3.dec g1964700 85 525
900917.2.dec 492415R6 1 469 33 902791.3.dec 3974703H1 86 3Θ2
900917.2.dec 492415H1 1 226 33 902791.3.dec 6051917J1 143 Θ15
900917.2.dec g12θ5162 8 427 33 902791.3.dec 6133494H1 172 469
90097.2.dec 33Θ520ΘH1 11 246 33 902791.3.dec 071200H1 204 383
900917.2.dec 3110432H1 46 322 33 902791.3.dec g76118θ 265 582
900917.2.dec 3469861 H1 99 369 33 902791.3.dec 3210822FΘ 422 948
999415.1.dec 2554389F6 1 428 33 902791.3.dec 3210822H1 423 603
999415.1.dec 2554389H1 1 251 33 902791.3.dec 423234ΘH2 501 748
999415.1.dec 5347358H1 2 261 33 902791.3.dec 1946681 H1 506 737
999415.1.dec 2554389TΘ 6 393 33 902791.3.dec 5078577H1 546 773
999415.1.dec 2733444H1 4Θ 287 33 902791.3.dec 2947767H1 554 858
999415.1.dec g1θ37277 159 376 33 902791.3.dec Θ3Θ4Θ45H1 656 946
999415.1.dec 2767114F6 281 702 33 902791.3.dec 1709770H1 688 909
999415.1.dec 579115ΘH1 380 682 33 902791.3.dec 450088RΘ 772 1110
900680.2.dec 2882704FΘ 1 490 33 902791.3.dec 450088R1 772 1365
900θ80.2.dec 55188Θ0H1 21 294 33 902791.3.dec 1420Θ07H1 824 1056
900θ80.2.dec 2Θ3959H1 53 361 33 902791.3.dec 3973901 H1 831 1142
900θ80.2.dec 2Θ99Θ7H1 55 367 33 902791.3.dec 4891459H1 950 1226
900680.2.dec gθ81529 5Θ 376 33 902791.3.dec 1Θ7Θ0Θ7H1 1012 1258
900β80.2.dec 2Θ3959RΘ 56 500 33 902791.3.dec 1Θ7Θ067F6 1012 1524
900680.2.dec 3385441 H1 57 304 33 902791.3.dec 337023ΘH1 1032 1157
900680.2.dec 6486823H1 62 645 33 902791.3.dec 2557147H1 1034 1272
900680.2.dec 4742910H1 85 350 33 902791.3.dec g2159741 1042 1437
900680.2.dec 4526902H1 90 362 33 902791.3.dec 2227Θ20H1 1123 1363
900θ80.2.dec gθ80873 325 609 33 902791.3.dec g4982966 1139 1595
900θ80.2.dec 1805Θ7RΘ 351 796 33 902791.3.dec 3527710H1 1183 1500
900680.2.dec 180718RΘ 417 796 33 902791.3.dec g2017626 1208 1549
900680.2.dec 1805Θ7H1 561 814 33 902791.3.dec ΘΘ3715H1 1207 1463
900θ80.2.dec 3742288H1 736 1012 33 902791.3.dec 1742349H1 1290 1585
900θ80.2.dec 2Θ7454H1 77 43Θ 33 902791.3.dec 4974283H1 1302 1572
900θ80.2.dec 2882704H1 2 260 33 902791.3.dec 4889859H1 1311 1593
900680.2.dec 180718H1 Θ88 780 33 902791.3.dec Θ2Θ7412H1 1359 1841
902791.3.dec g235072θ 1752 2098 33 902791.3.dec Θ599047H1 1397 1925
902791.3.dec g66θ724 1059 1324 33 902791.3.dec 2958222H1 1401 1714
902791.3.dec g712297 1903 2098 33 902791.3.dec 393523TΘ 1414 1796
902791.3.dec 4375518H1 782 1044 33 902791.3.dec 393523RΘ 1414 1842
902791.3.dec g1382485 173Θ 2168 33 902791.3.dec g5660784 140Θ 1856
902791.3.dec 3Θ5Θ99H1 1727 1863 33 902791.3.dec 3507311H1 1457 1750
902791.3.dec g5633687 1745 2168 33 902791.3.dec 458472ΘH1 1459 1750
902791.3.dec 1989220H1 1732 1934 33 902791.3.dec 1Θ7Θ0Θ7TΘ 1490 2106
902791.3.dec g21563θ5 1748 2170 33 902791.3.dec 450088F1 1497 2167
902791.3.dec g21θ5832 1752 2173 33 902791.3.dec g3038831 1521 1846
902791.3.dec 3728909H1 1771 2085 33 902791.3.dec g3038830 1527 1847
902791.3.dec g3178980 1788 2170 33 902791.3.dec g71229θ 1543 1785
902791.3.dec 4729789H1 1800 2064 33 902791.3.dec g75θ240 1543 1783
902791.3.dec 4858660FΘ 1792 2160 33 902791.3.dec 4829913H2 1550 1844
902791.3.dec g756133 1815 2145 33 902791.3.dec 18Θ3570H1 1551 1855
202 CD CO sf CD sf st cO sf oO O S sf - — CD CM sf sf OO st CD LO S O S OO S CM LO Ul Ul OO O lΩ O ω CN O CM CO CD S O CD st cO CM — CM CO Ol sf CM CN LO CM S CO CN st CN CM CN LO LO IO IO CD CO CD IO OI OO CN O O — — CD CD O O OO S S S CD LO O S CM sf S S CO CM — lO OO OO S — CM O CD OO CM sf CN — CN CO CN — — — CM CM OI CO CD — CM — CM CM CM O S C
— — — — — CO — — — CM CO C CO CO CO CO CO S CO CO CO CD CO CO CO st CO st CO CO CD st st LO st st CO CD LO S Ul st CD S UJ S CD S LO S S S S S S CD CD CD S S S S S S - O s
— — — — — CD — — — — — — — — T- — — — — — — — — T- — — — — — — — — — — — — — — — — — T- — — — — — — — — — — — — — — — — — — — — — — CN LO C
_
CD CD CD LO UI S OO S S — S CM sf OO CO CO O Ul CO CO - S S OO CD Ol CN LO S sf — — OI S CO CN S S S S CD CN O CO O O O — O CD 00 CM CD CN S S st S O OO O CD — S CN CM CO LO LO LO S OO CO σi σi — CN CM CO CO st sf Ul CO CO OO OO O O T- CO CO IΩ CD S S OO O O — CO CO CO CO sf Ul OO OO — — — — CM CM CM CO CO CD
© 00 00 CΛ O1 - — CM CN CO CD O O O O O O O O O O O — — — — — — — — — — — — CM CM CM CM CM CN CN CM CN CM CO CO CO CO CO CO CO CO CO CO CO st sf st st sf st st st st st — © OO OO OO OO CD CD CD CD OI OI - — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — —
Ul i-i u
o υ o o o o o o o o o o o o o o o o o o o o o o o o o o o υ o cvi CN cvi cvi cvi cvi cvi cvi cvi cvi cvi cvi cvi cvi cvi i cvi cvi cvi cvi cvi cvi vi cvi cvi cvi cvi cvi cvi cvi cvi cvi cvi cvi cvi cvi cvi cvi cvi cvi cvi cvi cvi cvi o o o o o o o o o o o o ooooooo o o o o o o o o o o
LO U1 U1 U1 U1 LO LO LO LO LO LO U1 U1 U1 U1 U1 LO U1 LO U5 U1 U1 U1 U1 LO U1 LO LO LO U5 U1 U1 U1 U1 U1 LO LO LO LO U1 U1 U1 U1 LO LO U1 U1 U1 U1 U1 LO LO U1 Λ CO CO CO CO CO CO CO CO CO CO CO CO CO CO C CO CO CO CO CO CO CO CO CO CO CO O CO CO CO CO CO CO C CO CO CO CO CO CO CO CO t
_u
-O
H
S O — CD CO CO CD CM OI CD O CD CO CM - OO S S OO OO sf cO CO S LO S CN S LO CN — 0O LO 0O OO sf O sf O sf — — l 01 — CO S O CD S S LO OO CO CO C
CD CD CD CN sf CO O CO O CD CN — CD O O CD CD CΩ — O Ol sf cO O OO Cvl Ol - CO CO S CO S O sf Ul Ul S CN CN - OO — CN CM — sf CM CO — LO IO CN O OO CΩ CD S CD CD — S S CD S CO C OI OI CD — CD CO — 0) CJ1 - CD CD CD O O — — — C0 CM O 00 O O L0 O O C0 O 00 O O sf C0 st - CD — CO LΩ S CD S S S S O OO CM CD Ul CO O O CD st S O O O — — — — — —
— — — CM — — CM — — CM — — — CVI CN CN CN CN OO — CM — CM CN — — CN CO CM — CM CM — — — — 00 — — — — — — — — — CD S — OO CO CD — — O1 00 CD — — — — — — — — —
— CO CO S S S OO — OO O O — CO OO S O — O S S S 00 00 CO sf CO 00 00 CO — sf CO — CO O s s CD LO S OO S S CD st st S CD CD S Ol O CO CO CO LO O S S st S lO — Ol O O S sf O Ol CN S CM CN CD LO S OO CD O — LO — CD CM CD CD sf CO S O O Ul CM CD OO - S S S S S lΩ LO - U LO lO LO CO CO CO CO CO CO CD S S S S CO O S LO CO LΩ CO CN CD sf CO lO S CO — O Ul CD CD S Ol st st st Ul Ul Ul LO LΩ S S S O CM CO st st LO CD Ol Ol Ul OO OO CO OO OO C
— — — — — — — — — — — — — — — — — — CD — — — — — — 00 — — — — — — — — CO CD CD CO CD — — — — — — CO CO CD CD CO S S S S S S S S S OO OO OO OO OO OO C
(D - CO —- — (Q - r - — — — — — CO — —- i— — CO *— CD U- I - H o)«r:i- i _- I I _- _- -: LO I I I I I I I I O S CO CM o — CC st x sM rI: Iz zI Xst X o — — CM usfιLO CD st i —N U -sfi —i UlcOβOrlioOisf O ι —i CO;lzι —oi —i CODCOCcsoficOLUolrI:cUNloOoiOoCDiCDiCDLU_lιUlιLO S O CD CM CD 00 CM CD O CO — O CD f CO CD CO LO CD OO CD Cvl st Ul T- CO LO O CD Ul CO LO vM st S S CM CM S O CD CD LO CΛ CN CD CD CD CO CO CO — S LO LO OI CD CD CO CD CO CD CO CO S — C LO CO Ul 00 S sf O st CD st 0000 CD CO CM CD CD st — S — CD CO CJl st CD S — S Ol CM S CD CM CO LO OO OO CO sf sf CO O OO st S S O sf sf st U CO CO OO CO O — Ul CM CD CN sf CO CO O C CO CD sf O LO Ul S S O CD 0000 sf S CD CD CO LO OO S OO — LO LO CO — — CM CD Ul OO CD S CM st S sf st OO CO CO CN Ul O OO CD S Ul CD CD CD CO OO sf CD CN — — O O CO st cO O LO CO C CO 00 LO — CD CM — — CM O CO 00 CM — sf LO OO CO st sf sf st O O CD sf CM LO st O st st cO CN CD OO st sf CD S CM - — — 00 — CD — S S S — CN O CO CD O O S CM — 00 00 01 00 00 0 OO sf U CM CM CO f CO CM CO O O CD Ol O CO CO CM OO CO OO S CO sf CD LO CM CD st sf CO st CD CD CN st Ul Ul sf st st sf st cO - CM O O st sf st LΩ O CO — — O CM CD — S S S S S S — CD COCO CO CO— CM CO CJ)— — sf — CO 0) 0) 01 CD— COLO CD— CM LO CD O — — CD— CO CD COCO — — CDCO CD LO CD CDsf COLO CD — — — S CO CM O O S CDCO sf CO LO LO Ul Ul Ul o o o o o o o o o o o o o o o υ o o o o o o o o o o o o o o o o o o o o o O O O O O υ o o o o o o o o o o o o o o o o o φ φ φ φ φ φ φ CD φ φ φ φ φ φ φ φ φ CD φ φ φ φ Φ Φ Φ Φ Φ Φ φ φ φ φ φ CD φ φ φ CD φ φ Φ Φ Φ CD φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ
XI XI XJ XJ XJ XJ XJ J TJ TJ XJ X) XJ XJ J XJ XI XJ TJ TJ TJ TJ TJ TJ x> XJ J XI TJ TJ TJ TJ TJ XJ J XI J TJ TJ TJ XJ J TJ J TJ TJ XJ XJ XI J TJ TJ TJ TJ XJ XI XJ TJ TJ
CO CO CO CO CO CO CO CO CO CO cb cb cb cb cb cb cb cb cb cb cό cό cb — — — — — sf t sf Sf sf sf st st st sf st st sf st st st st sf sf sf st st sf sf st st s
— - CO CD CD CD CD CM CN CM cvi cvi CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM CN CM CM CM CM C s CD s O Ol O) s Ol CD CD CD CD CD Ol Ol N CM CM s s C sD O O Ol CD sD CD CD D CD CM CM CM CM CM CO CO CO CO O CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO C
S S SSS sl s) CD sl s 01 O CD CD C S S S S S S S S S s 00 00 0000 00 CD CD CD Ol Ol Ol CD CD CD CD CD Ol Ol CD Ol CD CD Ol CD CD CD CD CD CD CD CD C
CN CM C sN s CN CN CN CN C CM CN CN CM CM CN CM CN CN CN sl C
CM CN CM CM CN CM CM CM CO CO CO CO CO st sf sf sf sf sf sf sf sf sf st sf sf sf sf sf st sf st st st sf st st st f s
O O O O O 0 O OOO O o O O O O O O O O O O O O O O O O LO LO LO LO Ul O O O O O O O O O O OOO OOO O O O O O O O O OO
CD CD Ol Ol Ol 01 Ol O O) Ol CD CD CD Ol CD O Ol CD CD CD CD CD CD CD CD CD CD CD O O O O O CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM C
CO CO C C CO CO CO CO C CO CO C C CO CO C CO CO CO CO C CO CO CO CO CO CO C C CO CO CO CO CO st st st sf sf st Ul LO LO Ul Ul Ul Ul U) Ul Ul Ul U) LO LO L^ cocococo cocoocoo cococococococo cococococococo cocococococococo co cocooocococococo coocococococo co cow
CM CO S CD CD CO CM S CO CD O 00 sf CM S LO LO — - Ul O C O O O O O LO S CO CD S O 0100 — LO CM sf CD S CO O O CO CM sf — st CD CO — - CM 00 S *^- o st sf o O CN CO LO CO O 00 CO O CO CD — - LO sf 00 00 CO 00 O O LO CD l 01 00 ^ S sf LO s o CN sf — - CD CD CD CD sf CD O O CM S 01 CD CD CD — - CN CN S CO
CM CM CN CM CM CM CM "" CN CO CM O CO CM CM CO CM O CM sf — CM CO CO CM CM CM CO CM O O CM CO CN CO CO CM CM CM Cvi O CN CO sf C CM CO CM CN CN CO O CM CO vo
__ o o LO LO CD CD S sf CO 00 CM CN CN CM CM CN CN st CM CM CM CM sf LO CO CO sf CO CO sf LO sf LO Ul LO S S CD S CD 00 CO 00 CD O •^ _. __ __ __ CN CD 00
© CN CM CM CM CM CM CN CM CM
Ul
H _. _. __ __ _ _ _ _ __ __ __ __ __ _ __
X x U — — XII — — T X —
X X — — ~r X X — —
CM CM I X CO o 00 X o CO X o x CM e I X X I — - CD LO M s s CD X o
CM , S — , — CM ^ — - CO LO sf U ^ CN — s CO ^5 CO CD ^ O CD t CO CM — - st CO LO ^f r — CO o
CD 00 o LO CO CO f LO Ol LO 00 O O 00 CD CO s o CM CN CN ^ — - ^t o ^~~ ^ ^ CO ^f ^t — - o 00 O o 00 Ol CO
CO r — CM CM f^. LO CO LO LO CD LO 00 I * CM — s s o CD CO o o s CD C5 CM 00 o o sf CO 00 ^
CO CO CD r*«. CD ^ s CD ^ CO CO ^" CD CO CO 00 sf sf 00 ^f CD — - oo CD CO CO CD O) CO — CD - — oo C_D CM st 1 — s 1 o o CD s s r
CM ,— ^ CO CO CD CD CO o CM s CD
CO st CO CD uϊ o CD ^ ^ O CD ^ CD CD CO CO ^- o O CD s CD st 00 CD sf CD CO 00 — - CO O CM — CD 00 oo st O sf st CO oo
-1- ,_ ^ ^ ,_ ,_ 1_ CO "* τ_ ,_ - Sf CM - CM CM CM ,_ CO O ,_ C |s" - CM CO 1_ "* CM ,_ ,_ CO CO CM CM CM CM ,— ,— CN "* τ" ,_ CM S s o o O o o o o o o o o o o o o o o o o o o o O O o o o o o O o o o ϋ o o o o o o o O O o o o O O O O o o O o o o o o O o φ φ Φ Φ φ φ φ φ φ φ Φ φ φ φ φ φ Φ φ φ φ φ φ φ φ Φ Φ φ φ φ φ φ Φ φ φ φ Φ φ φ Φ φ φ Φ Φ φ φ φ φ Φ Φ Φ Φ φ φ φ Φ φ φ φ φ φ Φ φ
-q -q -q -q q XJ q TJ -q q -q TJ q -q q q TJ -q -q TJ J TJ TJ TJ σ.-σ -q q TJ q q TJ -q -q q q q -q q -q q -ς q q -q -q -q ^j q XJ -q q TJ -ς -q -q -q q -ς q - s s s S s s s s s s s s s s s s s s s s S S s s s s s s s r^ s s s s s s s s r^ s s s s s s s s s s s s s s s s s s s S s s s s s s s
CO 00 00 00 00 O00 00 CO 00 OO CO 00 00 00 CO cό oό CO CO 00 CO 00 CO 00 cb oό CO oό 00 oό CO CO oό cό oό oό oό oό cb oό oό oό oό oό 00 oό oό oό oό CO oό oό oό 00 oό oό 00 oό oό oό o sf f sf st t st f st sf st sf sf sf sf sf f st st sf sf st f sf sf sf sf st sf sf sf sf sf sf sf sf sf sf sf sf sf sf sf f st st sf sf sf sf sf st st t st sf st t sf f sf sf sf s
CM CM CM CM CN CN CN CM CN CM CM CN CM CM CN CN CN CN CN CM CM CM CM CM CM CM CN M CM CM CM CM C CM CM CN CM CM CM CM CM CM CN CM CN CM CM CN CM CM CN CM CM CN CM CN C CM st sf sf st st st st sf sf sf st sf sf st st st sf st st st sf sf sf sf sf st st st sf sf sf sf f sf st st st ^ sf sf sf sf sf sf st sf f f sf sf st sf sf sf st sf st sf sf st st sf s sf st st sf sf sf st st st st sf st sf st sf st st sf f st Sf t st sf sf sf st sf ^ sf sf st st st sf sf sf ^ ^ st sf sf sf sf sf sf sf sf sf sf f sf sf sf sf st st st st sf sf sf s sf st st sf sf sf st sf sf sf sf sf st st st sf sf sf sf sf sf Sf sf sf sf sf st st f ^ st sf sf sf sf sf st st ^" sf sf sf sf st sf sf st sf sf sf f st sf sf st sf st sf sf sf sf sf st s ssssssssssssssssssssssssssssssssssssssssssssssssssssssssssssssssss cocococococococococococococococococococococococococococococococoococococococococococo
H l st O OO S sf S CO S LO Ul CD S CO CM O
LO OO CO sf st CM O O CN CD — Ovl OO st CN CN S CN — — — CM sf CD CO OO CO O CO OO CO S O sf LO OO st CM CD S LO LO st sf Ul CN CD CO OO O — O O CO CM OO CO CO CD CD O CO UI CO U CM O st cO CD CO st S O sf O sf S CO CO — OO LO S OO O CO OO O sf CD CO CO LO st — CO CM — CD sf S O — Sf CO CN — CN sf st LO CO CO — — LO CN CN S CM CM S CD LO sf lO st O — CD CO CD C
— CO st lO lO CO LΩ LO σi lO LO vO sr cO CO CD CD CO S CD LO — S CD S S OO S CD OO — CD — — CD — — — — — — — — — — — — CO CO CM — UI CO CO CD — CN CN CN CO CM CN CO CN CO CN C
00 — CO CD CD sf st st — st CN CD st S S OO CO CO LO CD CO CN CO CN st CD CO — — — S LO CO S S CD sf — C0 CD — CN CO st S CD
— LO OO sf sf sf sf Ol CM CN Ul CO S S O O O O — — — LO CD O O sf LO O st Ul O sf O - CO CD CO — CD CD sf oO CM CM CO CO st lΩ S CN Ol OO CD st — 00 CM S sf S C
— S — CN CN CN CN CN CO CO CO CO OO CO st st sf st st st st sf st Ul lO lΩ lΩ CO CO CD — sf — — I S OO — 00 CO O1 O1 — i- — — — CM CVI CD CO IO LO CO CO — CD 00 CD O1 CM - CO CM U1 — C
, — , — , — — - , — , — — — - , — — - , — — - — - — - CD — - — - , —
IIIIXLLlIUlIIIIXIIIIlXIr: —3 ZZ HI — I O I I oo I I I I I I T- X X X X X X — I LL — - I I I X I
O 00 CD O — - — - st 00 CO O S o CO CD O LO S S X CO OO st CO S S CD Ol o CO I CO CO I X CO CO O) Is- CO sf s CVI I CD CD I CM CO
Ul s CD Ul LO LO S S s f st st CM CO s sf s CD 00 st CO CO sf o CD CD CD CO CD CVJ , — S CM O CM CO CD sf r-
CD — S s 00 00 CO sf sf CD VI CO CO st CO 00 CO s s CM CO Ul CO LO — O O s o — - CD O s s σ> l o o Is- O O CO O
CO sf o LO st O O CO o> — - CN sf sf CN LO O t — - , — sf CD CO CM CM — CO CM s S CN o o I— o LO Ul LO 00 ^* , — CO CO
CO s CD O CO CD Ol CM — - LO st CO CO sf CM sf st 00 st s S CN LO sf S LO LO CO sf f^. ^ f^ r^* ^ s Ul CM S s O
Ol 00 s CN CM CO CD CD CD LO sf st CD S CO CN CO t O CM CO LO O sf O sf CD o CN LO r**» co O) O) CO ^ st CD o o s st CM st sf CN CM LO CM CN CO sf LO — s O CO CN CO CO CD O sf CO
" " o o O O o O o O O ϋ ϋ ϋ O o o " " o o O O O O O O O o o O O O O o o o o O o o φ φ Φ Φ φ Φ Φ φ Φ Φ φ φ φ φ φ Φ φ Φ Φ φ o o o o vo φ φ Φ Φ Φ Φ Φ Φ Φ φ φ Φ Φ Φ Φ φ φ φ φ Φ Φ Φ Φ φ φ XI XJ J J J TJ TJ q - -q -q -q TJ XJ TJ Ό σ TJ TJ TJ XJ J φ φ φ Φ o -q-q -q q -q -q -q XJ q a. -q -q q - -q -q-q q -q -ς q q q CD- CD oi oi oi oi oi σ> σ> σ> oi oi CD oi oi o oS oS oS σ> oi σi oi d ci CD- σi oi q p -q -q q - so s s s S s s
CN CN CM CM CM CM CM CM CM CM CM CN CM CM CM CN CM C CM CM CM CM CM s S 1 — s s s ι^ h* h* h- r- s s s ι^ r^ | — s Is- s oo 00 CO 00 0 O CO CO CO O CO CO CO CO CO CO CO CO O CO CO CO CO CO CO CO CO O o o o o o o o o o o oo o o o o o o o o o o o o o o o o o o st st sf st sf s
CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD 01 CD CD CD CD 0) CD CD CD CD O CO CD CD CD CD CD CD CD O CO CD CD CD CM CM CM CM
© sf sf sf st st sf sf sf sf sf st sf st st st sf sf st sf t sf sf sf sf sf st o o o o O o o o o o o o o o o o o Q Q o O O o o o o o Q o o ^ sf sf ^ sf ^ o O O O O O O O O O O O O O O O O O O O O O O O O O O o o o o o o O o O O O 3 o o o o o o o o o O O o o o o o o o o ^ sf sf ^ sf
CM ^
CM CN CM CM CM CM CM CM CM CN CM CM CM CM CM CM CM CN CM CM CM CM st sf sf sf sf st st sf sf ^* ^* ^* ^" ^* sf sf sf st sf ^* f st ^" sf sf ^" sf
Ul Ul LO LO LO LO Ul Ul Ul Ul LO Ul lΛ LO LO Ul Ul Ul Ul Ul LO LO LΩ Ul in Ul Ul Ul Ul LO CD CO CD CD CD CD CD CD CO CO CD CD CD CD CD CD CO CD CO CD C^ CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CTI CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CT
Ul OO OO OO S S O S CO O CO CO CO CO OO CD OO - OO Ul Ul O O Ul st CO O sf CM CD O OO O CO CO sf O sf O st S CO OO - CD LO LO L CD vD CO CO CD CO CO CD st O sf sf S S CD O CD O CM O st — — — — — — — — — CO CM CM CM CO CO CO CM CM CM CO — CO CO CO CO CM CO st st CM st cO CO CO st C CO O sf T- CD CD CD OO CD O - sf S Ol CD CD Ol CD st st st sf sf st st st st CN LO CD S CO CO CO CO S OO CD O CVI CN st S CD O CD OO OO — CO OO C CO st sf sf st st st sf st st sf Ul LO LO Ul CO CO S S S S CO OO OO O
II'- CO I I I I co I I I sf — X CM COOlsf CO
C sD i I _Z I I CO — co l s — co — sf s co σi s iooi o
CD CO O OI O LO CO LO LO — 01 f S CD CM
T- S CD — CN UI O UI OO LO S CD CD CVI CD C o S S O st S CD CD LO CO CO CD CO S CD CN C
01 sf OO sf
— 00 O sf O O - Cvl st C CM LO LO st cO C ,~ CD CM sf S CO CVl CD C0 O1CM CD CD CD CD CD C o o o O O O o O O O O O O 0 O O O O O O O O O O o O O O O O O O O 0 O O φ φ Φ Φ Φ φ φ φ Φ Φ φ φ Φ Φ Φ Φ Φ φ φ CD Φ Φ Φ Φ Φ Φ Φ CD φ φ Φ Φ CD φ Φ Φ CD φ Φ φ
TJ XJ J TJ TJ TJ TJ TJ TJ XJ TJ TJ TJ TJ TJ TJ TJ TJ XI J TJ J TJ TJ TJ TJ TJ XJ TJ TJ TJ TJ TJ XI XJ XI TJ J J TJ T st sf st sf sf st st" st sf- t sf sf' sf sf st st st st sf sf sf sf st f st st sf sf t sf' sf sf st sf st sf sf sf st sf s CM CM CM CM CM CM CM CM CM CM CM CM C CM CM CM CM CM CM CVI CVI CM CM CM CM CM CM CVI CM C CD CO CD CO CD CD CO CD CD CD CD CO CO CO CO CO CD CD CD CD CD CD CD CO CD CD CO CO CD C O) CD CD CD CD CD O O CD CD CD 01 01 Ol 01 Ol CD CD CD CD CD CD CD 01 Ol O CD C CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO C st st st st sf st st st st st st st st st sf st st st st st st sf sf st st st st st st st sf sf sf st st st st sf st st st sf st sf st st st st st sf st st st st st sf sf st sf st st st sf sf s
O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O sf sf st st st st sf sf st st st st st st st sf sf sf sf st st st st st sf sf st sf sf st st sf sf st st st st st st st st sf st sf st st st sf sf sf st sf sf st st sf st st st sf st sf st sf st st s
H
LO sf sf CO CO CO CO LO CM CD CD T- CD C
O sf CN CD sf Ul - CO st — CD CO — CO OO CM CO sf S CD CO LO — CO O CM CN S IO OO S CN OO CO CD CO CO S IO CN CO CO CN U O — CO CD CN IO CD CD S O CN CD U U UI UI O CO CO CM CO S CVJ O st CO S LO — CD CO CO CN LO CD CM CO OO IO CO S OO S CO CO CO O O CD OI CO CN — S — CN CD — — S CD CO — CO S LO CD O CO CO — sf O O st st Ul O Ul LO LO LO CN CM CM st CO CN U CO LO CO CO CM CO CO CM CM CM — CM CM sf - CN CN CM CM CN CM CM Cvl CO st sf CO CO CO st st sf st st st sf sf st — CM CVI CM CM — CM 0O S 0O S S 0O — Ol CD Ol — — — — — — — — — — —
CM CM O CD OO CCOO ULOl OO CC
CD CD CO CD CO CD — LO CO LO CD CD S CN UI OO UI CD OI CO O O CO LO L — CD CD LO S CD sf S CD CD OO S S — CO CN O LO st CO CM CM CO st S OO αi O — — C CN sf CD i r-- CCDO SS CCOO —— LLOO SS OO —— C CMVJ CCMM OO ssft ssft sf CO Cvl CVl CO CO LO S O O O O CO st OO — — CO — st — — — — CD — — CM CM CM CO CO LO S — — — — — — — CM CM CM CM C CM CM — CM C COO CCOO CCOO ULOl CCOD ssff ssff LLOO LLOO LLOO UUll SS SS SS S CO - — — — CD CD — — — —
I H I I I I I I
s
— - , —
S S S S S S S S S S S CO CO OO OO CO C0 03 00 00 CO CO OO OO M OO CO CO CO CO CO CO CO OO CO CO CO CO CD CD CD O) <D 0101 01 01 01 CD CD CD CD 01 01 CD » CO CO CO CO CO CO CO CO CO CO CO CO C CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO M
st CM S S sf CM S O Ol CO — O S sf CD S LO O sf sf S CN Ul — S O OO CD CD OO CO S CO — O CO CD CD sf CD CO sf OO st S CM CN O - CD sf CM sf S S - — O CM CO CD sf O S Ul — 0 st CM OO OO Ul LO CO O O CO CD st CD CM Ul OO CD CO Ol Ul Ul CO — st st cO CD CO LO OO CM st CO LO Ul — S S sf cO S CM O OO CO OO CD LO CD O CD — — S CD S sf CD CO OO OO CD st LO CO CO C st cO Ul OO CD OO S OO OO CD O - — CN — CM CM CN CN st CM CO CO sf sf CO st CO st S Ul CO st st CD S CO S S O OO CD CN CM — — CN 00 O O CM — — O Ol — — CVI CO O O O O CD CO CO —
— — — — — — — — — — CM CN CN CN (M CN CN CM CM CM CM CM CN CN CM CN CN CM CM (M CM CM CM CN CM CM <M CN CN CO CM CM C C CO CO CO CN vo
IΛ S S st CD CD st sf S S O CO sf - S CO CD OO CD CD CD CD CO CD — Ul sf CO sf Ul lΩ Ul CO OO CD - CD — CD S CD — — sf oO CD - S S CO S S CO CO CD CM O CO OO Ol CD S S S S CM S sf CO CO CD S S — — S CM S CO CO sf LO Ul CD CD - — CD CD O CO sf CD CD O O O S S CO CD CD CO CO CO CO Ul Ul CO st CO OO O CO st Ul S S CD O O O - LO LO UI UI CD CO CD CO CD CD CD
© O O CO CO CO sf st LO CD CD OO OO CD CD CD CD CD O O O O - — — — — CM CN CM CN CN CN CN CM st st sf LO Ul Ul CD CD CO CD S S S S S S S OO OO OO OO CO CO OO OO CO CO OO OO OO OO CO C © — — — — — — — — — — — — — — — — — (M CN CM CM CM CM CM CN CN CN CN CN CN CM CM CN CM CM CN CN CN CN CN CN CN CM CM CN CN CM CN CM CN CM CN CN W Ul
H cor: co ϊ — CD "*— *— CO *— CD '— i— '— •— '— CD CO 1— ' ' CO '- — — — — -— CO — — — — — — U LL I CO LL I II I sf X X LL X r- I Li. I I I I I li. X — X XI I IH U O HII I H I I I I IIH -Il l l sf u
CD CD O CM CM- Ulϊ Ulϊ CDϊ CD LO CO CD — — CM C CD O CO CO 00 I 00 CO CO — S CO LO CD CD O CO 00 CO 01 CO ICO IO I— ICO IUl st CM τ- CO CD CO — OI OO O S CD OO CD S LO O S CO CO CO CO CD CD
PH CO CO sf — — S S CO — CD O Ul CO CM CM S CN sf CO 0000 CD CD Ol sf CO sf 00 S 00 00 O CO CD 00 S — O CO st — sf CD — sf oO OO S st CM OO LO CO CO CD CO S S CD IO S CVl S CN CO CO C
O O LO CD CD CO CO CO Ul CO O CD CD CD CD CD S O 0 CM CM CO CO O S 00 CO CD S — — CD CD O CO Ul Ul — S sf CM st O CΩ IΩ LΩ — OO S CO CM CM CVl OO CO Ol CO CO sf st O CM CM O CO - f sf CO sf st — — σ> s 00 LO CM CM — LO CO CO 01 01 S S CD st sf sf 00 S LO O 00 OO CM CO — sf Ul O 00 O CD CD Ul CD st OO σi OO lO CN CD S O O CN CO CO CN — LO CO CN CM O O S —
CO CD CO sf sf — — CM CD S S 00 LO 0100 LO CD sf sf — — 00 — sf LO CO U O CO 00 00 CD — CD O O CD sf CO O CD sf CO sf OO CN OO sf CO Cvl - — — sf O CD sf - — S OO OO OO OO CO sf L
CO CO CN CD CD CO CO CM sf — l CO — S S CD — sf sf O O sf CD CN S CM 00 CO CO CM CM — — CD sf S CD 00 — sf U — CM CD — — CM OO S sf O CO CO CD CM CN S CO CO st st sf st st sf CO
CM CM CDCO CO sf sf — LO sf CDCD CM CDsf O CDLO Ul st st CO — st sf LO LO Ul sf — — COLO COUl CN sf — sf CD CM sf CD CO CD CD— LO CO CVI sf sf st sf st CO - sf CDLO U1 U1 U1 U1 CD OIC o o o o o o o o o o o o o o o υ υ o o o o o o o o υ o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o υ o o o o o o o o o φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ
XJ XI TI XI TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ XJ XJ XJ TI TJ TJ T^ st sf st st^ st st st sf st sf st st st sf st ^ st st st st s^ CD σi oS cD CD CD CD CD CD σi σJ CD CD CD CD CD CD CD CD CD CD CD CD CD σi Ol Ol Ol Ol Ol CD CD CD oo oo co co ∞ co oo oo co oo co oo co co co co oo co oo co co oo co αs oo oo rø
CM CM CM CM CM IΛl CVI CM CM CM CM CM CM CM CM CM CM Cvl (N CM CM CM CM CM CM CN eM CM CM (M CM CM CN CM CM CM CM CM W O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O αo w co co co co co co co co co co co co co co co co co co co co co co co co co co co co co cr> cr> co co co co co w
st st sf st sr st st sf st st sf st st sf st st sf st st sf sf st st st st st st sf st st st st st st sf st st st st sf st st st st sf st sf st st st st st s^ t
3
— sf CO CO OO lO S CD sf CD Ul S O CO CO CO OO S CM O I st CD sf st O CO OO OO CD O O CM - OO S CD Ul OO OO st CO st cO S CO Ul S S st CD CO CO CO CO - O U1 — CD OO sf CO sf CD CM CN CD sf sf st CO CO CO CO LO S CO CM CM O S CD OO st CD CM U O S CO S CD OO OO LO S CJI CD CO — CO CO CD CD st O CM Ol LO LO CD CO O CN S - Ul S CO sf — — — CN CN — CM — — OO CD CO LO LO LO CN — LO S O O — — CM C st CO CO CO CO CO CO st cO CO CO st Ul Ul Ul CD CO LO CD st S CD S CD S OO OO S CO OO OO CO OO — 00 — CD CD CD — — — — — — — — — — — — — — CO S LO Ul LO LΩ sf Ul CD — — — —
O CO OO — CO OO CD CM CD CD CD CO — 00 O S S CD — CO UI U — — — sf O - S OO sf CN OO S CO S CD CD S Ol CO CO CM S S CO sf st cO CO - 00 — — — — — — CD LO CO OO sf
O O O CO CO CO CO CO O O — CO LO S OO OO OO O st st cO OO st sf sf CO S S OO OO CD O O — CO CO S S CO O CM sf S — — CM CVl CM CO sf Ul — O — OO OO OO OO OO - CM CM UI CO UI
— — — — — — — — CN CN CN CM CM CN CM CM CM CO CO CO st st LO Ul Ul Ul Ul Ul Ul Ul Ul CO CO CD CD CD CD CD S CO OO OO OO CD CD CD CD Ol Oi σi CD - CM LO CM CM CM CM CN CM st OO OO CO CD CD C
stooll — IcoII i LO l I I I st o I I co I — ir CD — X X I I I I ii I _r I I I CD - I CCI Il-I I st — III I II I I XII I I sf Ul CO O Ul sf CD S CD O CD CN OO CO CO O CD OO LO — S CD I X I CD Ol LO Ul CM S — 00 CO I LO 0I00CD0 Xst CID sIf LO T- sf U CM LΩ LΩ CO l lO — — OO CM Ul — S CO CD CN CO CO — CO S CM S S — S CO CD sf - l CO Oi σi CN O CD sf CD st CM CM sf O sf CN sf CVI O CM sf CD CO CO CD CO 00 CO — CO CO CO CD 01 S CO CO CM — OO st — sf CN CO S CO Ol — CN CO CO O CO CD CN sf o — coσioσicocD CVl CD O LO CM OO CO sf CN Ol O sf O sf S O CD CD — CM CO Ul CO LO Ol CD — — CO CD 00 — 00 CO — — — S Ol O Ol S Ol O O O sf LO CN CD LO — CO CD CO S
OO CD OO LO sf - sf cO CM cocD co σi ui — oo oo s CO O — CO sf sf S CO LO CD S CO sf CO CD CD sf CD CD CD O 0000 CO CO Ul — — CO O O — O CO — — — CO CO - CM — CO O O CD sf S CO — 00 — O — CD CD sf O st CD CO S - — sf sf sf CD CD sf CO — Ul CD CO LO — CD CM CD U O 0000 st LO CM O CM CM sf CD CD — — — O — LO CD CO OO S S OO CD CD CD — CO CD sf — — LO CD CO CO CO CO CM S CO — CO CO OO CM CM CD CD — CD CN S S S O LO sf O O LO U CD Ul CD — CD 01 LO sf sf — — CO CN CM S CD CM — st sf O st st sf sf sf CO O O S CO CO CO CD CDCD CN CDCO C CO LO sf C CO sf sf — C vJ — sf CDCM CD- — sf CM — CO sf LO U CD O CD CVI CD— LO Ul CN CO CO COCD CM CM CO — Ul st CD COCO lO LO LO LO LO LO LO LO sf cO CD L
O O O O O O O O O O O O O O O 0 O O O O O O 0 o O O O O O O O O O O O O O O O O φ φ Φ Φ φ φ φ φ φ φ φ Φ Φ Φ Φ Φ Φ CD φ o o o o o φ Φ Φ Φ Φ Φ φ φ Φ Φ Φ Φ φ Φ o O φ Φ φ φ oφ o Φ φ φ Φ Φ φ φ Φ Φ φ Φ Φ Φ O O O O O O O O O O O O O O O O vo J TJ J TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ XJ XI TJ TJ XI XJ J J TJ TJ TJ TJ TJ TJ TJ XI XI TJ XJ X) XI TJ TJ TJ TJ TJ TJ TJ J XJ X) XI TJ TJ TJ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ o f sf f sf sf t st sf st sf st sf sf st sf sf sf sf sf sf sf sf sf st sf st st st sf sf- sf sf sf sf sf sf sf sf sf sf sf st sf sf sf sf st sf st sf sf' XJ TJ XI XJ XJ TJ TJ TJ XJ XJ XJ XΪ Tl XI Xi TJ so CM CM CM CM CM CN CM CN CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM CN CM CM CM CM CM CM CM CM CM CM CM CM CM CN CN CM CM CM CM CM CM CM CM CM C CM CM st st st sf- st st sf sf st sf sf st st st' st- st
CD CD O CO CO CD CD CO CD cb cb CD CD CD CO CO CD CD CO CD CD CD CD CD CD CO CD CD CO CD CD CD CD CD CD CD CD CD CD CD CD CO CO CD CD CD CD CD CO CD CD CD CD CD CD CD OI OI OI CD CD CD CD CD CD CD CD
CD CD CD CD CD CD oi oi σ) CD CD CD CD CD CD CD CD 01 Ol Ol CD CD CD CD CD 0101 CD CD CD Ol CD CD CD CD CD CD CD CD CD CD CD CD CD Ol 01 CD CD CD CD CD O CO CO CO CO CO co o co CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO O CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO 0O 0O OO OO CO CO O0 CO O0 CO CO 0O O0 0O CO 0O
© CM CN CM CN CN Cvl CN Cvl CN CN CM CM CM CN CN CN o O O O O O O O O O O O O O O O O st st st st st sf sf sf sf sf st st st st st sf st st sf st st sf sf st st sf st st st st st st st st sf st st st sf st st st st sf st st sf st st st sf co co co co co co co co co co co co co co co co
OOO OOOOOOOOOO OOOOOOOOOOOO O OOO OOOOOOOOOOOOOOOOOOOOOO — —— — — — — — — — — — — — — t st st st st st st sf st st st st st st st st st st st st st st st sf st sf sf st sf sf st st st st sf st st st st sf st st sf st st st st st sf sf sf st
S LO CO — S CO — sf CM CO sf O CO sf S CO —
LO S S CO CD OO CVI CO sf CD CN OO st lO sf CD OO LO — S 00 sf τ- 00 O O CM O CD CD CD — CM — CD O CD S OI OO CD O - sf - CD CM CM C
— Ol O sf st cO CD Ul CO CO S CO OO - 0000 — — CD O OO CM S LO O CO CN — S O0 — S OO CO st CN CM CO CD O CO CO S CM O S - S S — CO LO CN CO sf — S sf CD CD CM S O LO O O
— 01 - — — — — — — — — — LO CO — CO S OO CO CD CO st CO CO st st CO CO CM S CO CM CO CO Cvl st sf st cO CD CD st LO CD S S CD CD OO - — CM CM st st Ul S CD CD CO — 01 — CD — — vo
IΛ sf CO - — — S CD UI O O O O CO Ol st S S S CM OO Ul CM CM S CM CD Ol CN - CO O CO CN LO CO CO CM CO O CM LO LO CO CO — sf LΩ CD OO OO CD CD S S S S CD CD O
© CO CO OO CM CM CO CO CO O — — CN O O CN UI CD - O LO S CO CM CM LO Ul sf CD OO OO CD CO OO S S CD CD CM O Cvl CD - Ul OO CO S S OO CO O O CM CN sf Ul © CD CD CD — — — — — — — — — — — — sf sf Ul - CO — 0000 000000 00 — CM CO CD CO OO — — — — CO — — CO CO IO IO CO S S CO — — CM — — CM CO CO CO CD S S S S S S Ul
H — — —- — U —- —- CO CO r_ — — — —-
CD
_I_ I I CD III —- H LL LL Ul X I I I I CO CO I — CO — stcθCDlIlLL.Ili.IIIIlL_IIIIIIIIIII — LOllL-IIIU-IIII
H o CD st S L S s s CD S Ul st s Ul CD S CD CO CD S O CD O sf sf LO CO CM CM sf OO O CN CO —- ,— CD LO S LO sf S 00 CO CO CM I CD O OO CD O) LO l CO CO —- —- I
00 ,— sf — C
CO CO CO CD st CO LO sf S CO s s s S Ul 00 CM s Ul l CO O CD 00 CD CM OO sf S 00 00 CD CD CO CO 00 CD CN s s s sf O S sf s CD CD CO l 0000 s O CO sf CM 01 01 CO 00 s st s
CD CD CD CN CD CD O O CM Ul CD CO CO CO oo s st S CO CD — CO CD 0000 — CO — CO CD —- O S LO l CM —- sf O CO st sf s Ul CO —- ,— S LO S CM 00 CD LO CO CO CD Ul o o s CM CO 00 C
Ul Ul LO CD 01 LO CD CD st 00 IO s s s l CO sf s CO CO CO CO 00 S O CO O CD sf Ul CO CD CM CM LO CM CO S st CD CD CO CO CM sf CD CD sf st CD O sf CM CD CD S 00 CM CM CO ,— CO CD 0
O O O O CM O CD CD S CN O CD CD CD CO CO CN CM CD s CM sf S 00 — CO f sf — O CD sf sf CO CO CO O sf st CM st LO CD CO —- CM s ,— CO o O CO CN CM O O LO LO CO CO 00 LO C
00 CO CO O CD OO st st CD CO 00 CN CN CM CN U LO CD CN ,— LO sf sf sf LO sf sf CO sf sf LO CD CO CD CD CD CO CO IO sf O CO CO s LO LO —- LO st s CM st 00 CM CO CD 00 CO 00 O CO 00 o —- sf CD C τ" CM '- '- Ul Ul LO CO — COUl LO LO CO CO CD CD O O) CJ) CJ) O) CO O '- *- CM CM CO LO Ul Ul CD LO LO CD O CD CD CD CO CD CO CM LO COCO LO CM CM CO sf CM CM CO — C o o o O O o o O O o o o o o o o o O o o O O O O O O O o o o o o O O O O O O O O O O υ O O O o o o o o o o o O O O O o u o o o o φ φ φ Φ Φ φ φ Φ Φ φ φ φ φ φ CD φ φ Φ φ φ Φ Φ Φ φ φ φ φ φ φ φ φ φ φ Φ Φ Φ Φ Φ Φ Φ φ φ Φ Φ Φ Φ Φ φ Φ φ Φ Φ φ φ φ Φ Φ Φ Φ Φ φ φ φ φ φ Φ
XI XI J J J XJ TJ TJ X) TJ XJ J XJ J XI XJ XI XI XJ XJ TJ TJ TJ TJ TJ TJ TJ J TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ T cb cb cb CO cb cb cb CO cb cb cb cb cb cb cb c cb cb cb cb CD CD CO cb cb cb cb cb cb cb CO cb cb cb CD cb cb cb CO cb CD CD CO CO CD CD CD cb CD CD CD ui ui ui ui LO¬ ui ui ui ui ui ui ui ib ui ib i ui ui ui ui ui ui ui ui ui ui ui ui ui ui ui ui ui ui ui ui ui ui ui ui ui ui ui ui ui ui ui ui ui ui ui ui ui ui ui ui ui ui ui ui ui ui ui ui ui ui ui CO b oό oό CO oό oό cό oό oό oό oό cb oό cό c
CM CM CM CM CM CM CM CM CN M CM CN CM CN CN CM CM CM M CM 01 CD CD oi σ> oi CD σi σ) oi CD CD CD 01 CD CD CD CD CD 01 CD CD CD Ol CD CD CD CD CD 01 CD CM CM CM CM CN CM CM CM CM CM CM CM CM CM CM C
LO LO l LO LO Ul LO LO LO LO LO LO LO LO LO Ul LO LO Ul Ul SSS s s s s s s s s s s s S s s s s s s s s s s s s s s s S CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD C
CD CD CD CD CD CD CD 01 01 CD CD Ol CD CD CD CD CD 01 CD CD CMCMCMCvlCM CMCNCN CNCM CM CN CM CN CN CN CN CM CM CN CN CM CM CM CM CM CM CN CM CM CN O O O O O O o O O O O o o o o
CD CD CD CD CD CD CD σ CD CD CD Ol CD CD CD CD 0) CD CD 01 CM CM CM CN CM CM CM CM CM CM CN CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM CN CM CM ,— —- —-
CO CO CO CO CO O CO CO O CO CO CO CO CO CO CO CO CO O CO CM CM CN CM CM CM CM CM CM CN CM CM CN CM CM CM CM CN CM CM CM CM CM CM CN CM CM CM CN CN CM sf sf st St sf sf st sf st sf sf st sf st st s
CO CO O CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO sf st sf sf sf sf sf st st st st st sf st st st st st st st st st st st sf st st st sf st sf lO Ul Ul Ul Ul Ul Ul Ul U^ st st st st st st st st st st sf sf st st st st sf sf st sf st st st st st sf st st st st sf st sf st st st st sf st st sf sf st sf st st st sf sf sf st sf st st sf sf st sf st sf sf sf st st sf sf s st
_-
X)
CO CD CO LO CM CO S S LO U CD — — sf CD OO sf st CD sf CO CN S CO LO - S O CO CD O CD CO O CVl S CO sf sf CO LO CO CD CO O — sf CD 00 O CO C
CO CM CN O CO sf CO CN CD CO CO st cO CO CO C CO CO CD CO CO — CO CO CO CO Ul — CM O O st CO S O CN CO CO CO lO CO st — CO OO LO Ul sf Ol S CD S CM Ul CD OO LO CD sf O CM CO S - LO CO L CO O — O CO CO CO CN CN CO CO CO OO CO CO CO CO CO CN CO CO CN CO CO CO CO — CD LO CD st st σi CD CO CO LO st CO CD S CN — O — — — CN CN st CO CO S CO CN O — — — S — CM sf sf CM CM U CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CD CO CVI st CD LO S CD - — — — — — CM CN CM CM CM CM CM — CD CN CO CN CO CO — — — — CD — — — — — —
CVI CN Ul Ul Ul sf OO CD CD CO CO - CM S — Ul LO LO CD S CO O CO sf Ul LO S oo o σi oo co co LO O o co co co — — — CN LO LO — CM CO sf sf Ul Ul CD CN — CO CO CO LO OO CD — CD CD CO S CO OO S sf sf CO - CO Ul Ul Ul st O CO CO CD Ol CD st O 00 S 00 CD C0 O O O O — U s σi σi σi oi o cD o o o o o o o o — CN CN CN CN CM CD CD O CD CD CD CD Ol CO CD S sf sf CD O - — sf sf CO S OO OO OO OO CD lO - S OI OO OO U S S S CD O — LΩ IO LO CO
CM CM CM CM CM CM CO CO CO CO CO CO CO CO CO CO CO CO CO CO CVI CN CO CM CM CM CN — — — CM LO LO CO — — — — — — — — — — — — S S — CO LO S OO OO OO OO CO CO CD O OI OI CD OI
CO J CD -: — — CO — — CD '- '— '" CO CO - co — — — — — — — T- T- - CD '- '— CM CD S *- I" L I I I 0100 CM I I — CO sf i cDst — CM = — I IIDC I I I_ IIII_-: Γ:H CDII_.I I I IIII_-I I -:I_ III r- — Li. U_ X X
O O sf sf CD CO S Ul CO — sf cO CD — CO O — — — CD — CM CD CD OO CD l Ol CO CO CO CO S — — CO S CO OO DC I — H O O O σ OO sf O CD O l S CO l CO CO CO CN I CM CD CD CD O L
LO LO CO CD 00 CM CD S S st σi sf cO CD S O LO lΩ lΩ st O CD CO O O O LO O CM CD CD CD CM CD CD CD CO st st CO CO CO LO S S — sf cO CD OO - S — sf CD OO CD CD CO - OO OO S S S S C
CD CD CD CO Ul O S Ul CD O S O O LO CO sf S S S O CM LO lO σi CD LO LO S CD Ul CD CD st Cvl CM CD CVl CO CO sf st CN O O O - C0 U1 CD CM LO C0 CN LΩ — lO CD CD — CO LO Ul LO LO LO CO 0
S S 00 CM CM O 00 O OO O CD LO — S CD CO CM CM CM S co σi oo σi oi co LO O s S S S Ul CO CO sf CN OO CO — — CO LO — — CD OO O O OO sf S O CO CN OO OO OO Ol sf S OO st sf sf S 0
Ul Ul O — CO sf CD sf CM — O LO CO CO CN LO CN O CO O sf CN CM S O O LO - CD CD CM CM O — T- — CD CD CD CN CD sf sf T- CO CN CM CM Cvl OO st Ul CD O CO CO LO 0
— — sf CO — CO sf CD CO sf sf oO CD LO - CD S S S — st OO sf sf sf — sf CN LO sf O O O sf st st Ul OO OO O O sf CO LO Ul LO LO CM — sf OO S sf CJl lO CD CD CD CM lO O — O O O — U
CN CM Ul CM CO CO CO— CM CD CO CD Ol O) CO CO"— — — CM CDUl CO CO CO COCN COCO sf C CN CO — — CD CO IO LO CD OI — Ul - — CM CM CD CO CM St CO CO LO Ul sf CM CM st Ul sf COCO CO CO CO C
O O o o O O O O O O O O O O O o o o o o O O o O O O o O O O C) O O O O o o o o o C) C) o o o o o o o o o o o o o o o O vo ω CD CD φ Φ Φ Φ φ φ φ φ φ φ Φ Φ Φ Φ Φ Φ Φ Φ Φ CDoΦoφ o Φ oΦ Φ Φ o Φ Φυ φ φ oφ oφ Φ Φ φ φ Φ φ φ φ φ φ φ φ φ φ Cl) φ φ φ φ φ φ φ φ φ φ φ φ CD φ φ Φ o TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ XI TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ XJ X) TJ XJ X) XJ XJ XI XJ J TJ TJ XJ XJ XJ XJ XJ XI J XI XJ XJ XI XJ J XI XI XI J J XI J XI XJ J J TJ TJ TJ TJ T so sf sf sf sf sf sf sf sf sf sf sf sf sf sf sf sf sf st st sf st sf st sf sf sf st CN CN CN CM CM CM CN CN cvi CN CN CM CN CM CM CM CM CM CM CO CO co co co co co co co co co co co co co co co cb c
CD CD CD CD CD CD CD CD CD CD Ol Ol Ol CD σi σi oi Ol CD CD σi CD oi o) σ) oi σi sf sf sf sf sf sf sf sf sf sf sfsf t sf sf sf sf sf sf sf sf LO LO U LO LO LO l LO l Ul LO LO LO LO LO ui u
CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CN CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM C CM C co oo oo oo 000000 CO 0000 CO 0000 00 CO CO CO 0000 CO CO 00 0000 oo CO 00 S S S S S S S S S S S s s S S S S S S SS U Ul Ul LO U l Ul LO Ul l Ul Ul LO Ul LO Ul Ul Ul
© CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM CN CM CM CM CM CN CM CM CM CM CM CM 0000 00 CO 0000 00 00 CO CO CO 00 00 CO 0000 0000 CO CO 00 CD CD CD Ol CD Ol CD CD CD CD CD CD 0101 CD CD CD CD C o O O O O OOO OOOOOO O O O O OOO O O OOO O O CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CD CD Ol O Ol Ol CD CD 01 CD CD CD Ol 01 01 CD CD CD C
CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO O CM CM CM CM CN CM CM CM CN CN CM CM CM CN CM CM CM CN CM CM CM CO CO CO CO CO CO CO CO O O CO O CO CO CO CO CO CO C
— — — — — — — — — — — — — — — — — — — — — — — — — — — CM CM CM CN CN CN CN CM CN C CN CM CM CM CN CM CM CM CM CN CN CO CO CO CO CO CO CO CO CO CO CO CO CO CO crj CO CO C st st st st st st st st st st st sf st st sf sf sf st st st sf st st sf st st st st st sf st st sf sf st st sf st st st st st st st st st st st st st s^
o o o o o o o o o υ o o o o o o cb cb cb cb cb cb cb cb cb c σi σi σi σi oi cji σi oi oi σi C - σi σi cb σj σi CD oi oi σi cji o o o o o o o o o o o o o o o o o o o o o o o o o
CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CD CO CO CO Φ st sf sf st st st sf st st st st st st st st sf st st st st st st st st sf st st sf st st st st st st st st st st sf st st st st sf st st st sf st sf st st sf st st st sf st sf sf sf st sf st sf st s t
J-
X)
E-i
CO CD S CO S O CD CN — — CO CN OO lΩ O st LO S lO OO lO O CD CO O S — LO — σ CM CM LO CO S CO CO CO st — CD S S st cO CO CO OO — CD CO CD sf LO — S CM — CD sf — O CO sf CD CD s sf Ul Ol lO CD LO sf st OO CD S CO — S S — O O CO O CD OO CD lO st — S S CO CN CN sf — — CM CD CO sf CO CO Ul CO CO O OO st CO CO sf CO S CM CO CN CD CD CM sf CM CO CO lO CO lO CM CD C
— O — CM — CN CN — CO LO — st sf Ul st st oO Ul sf oO st lO Ul CO eO ω Ul CD OO OO OO CD OO OO O S S CD OO OO CO CO CO — OO σi CM — — — OI OO — O OI CD O — — O O O — o o o
— — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — CM — — — — — — — — CM — — CN LO lO LO st st Ul LO st sf lO LΩ LO LO LO LO LO LO LO lΩ L
CO OO O OO CO OO O CVJ O O — U sf — CO CD CD S CO CO sf S OO CD LO LO — CO — CM CM CO LO OO st sf CN OO — CO CD O LO S OO S — 00 O 01 O S — O CO CO IO O CO CD LO L OO Ol S sf Ul CM sf lO OO - — CO LO O O O O — CO CO CD OO sf S S S S S CD lO CO CD CD — CM CM CM LO CO S S Ol CO sf st S st — — — — CN CM sf st lO S CO OO OO OO CD O - — CO C O O CM CO UI O O O O — — — — CN CM CM CM CM CN CM CM CM CO CO CO CO CO CO CO sf st sf st Ul Ul Ul Ul Ul Ul Ul Ul Ul CD CD CD CD OO S S S S S S S S S S S S S S S OO CO OO OO O CD OI CD CD CD - — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — st sf st st st sf st st sf st sf st st st st st st st st s
— - co *- >- — CD CO CD _ _- — O s-
I r: I co I I I I I I I I cc I I -> I I H ΓI I I I Ul I I I CD CN H st H S li. I I s σil ui o cD ui coI Ol sf S st I I co I r: ΣZ. I CO H I f I LO I I I
Ul I CM — CD — CD — LO sf 00 CO sf O st CD I CO Ol CO S CD Ol 00 CO CD 00 CM — CO CO CD — CM — CM CO sf CM CO CO lΩ CM — CD f Ul O CM — I I CO CO CD sf S CM LO — CD CO C
S CD O sf — O S LO CO CO s 00 CO CO s Ol CO U — sf — O 0000 T- CD CO CO st — — s S — CD CO — CO Ul σi CD — s st sf sf O S 00 — sf CM S O OO LO O OO — CD O CD L st S sf CD CO S O CO CD CM 0 sf CM sf — CO CO LO CO Ol CN CD S 00 CN CN O sf O LO S S sf CD CN CD — CN LO O UI CO CO O CN OO Ul O S CD — S S S CO S CO CD CO CM CD
CO CD S CO O CO 0000 CD Ul s sf CN LO CM O 01 CO CO sf CD CO CN CM CM CM CO sf CO O LO sf sf CO CO CO OO CO O LO — lO lO O sf LO CO CO 00 O LO CO LO S CM CD Ul Ol CN CD — CO 0
CD S OO CD — CO OO CN Ul sf o CD LO sf Ul CM CM — CO S CO CM — O CO CO sf CD S — CD O O CO S CO CO CO O CM CO CD CO S — S S CN IO CD CO S CD 00 CD CD CD CO CO — LO CD
00 S — — Ol sf CD CD OO O 01 CO O S S 00 Ol S O CM LO st Ul CD CM CM CD st S — CD CD CD CN lO S sf CN CM sf — O sf CM CVI CO 00 — sf LO O 00 CM LO CO — O sf — CO LO C
CO O CM CO— CM — CM IO CO — — — — CO CO sf CM Ul CO CM CO COCO sf CO CO COCO— COCM CM LO C0 CDC CD O) O) C0 CDsf C) CO CO CO sf CD COCO CD CD CM CO— — CDCO CDCO CM sf s o o o o o o o o o O O O o o o o o o o o o o o o o o o o o o o o o o o o o o vo φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ o X) XJ TJ J so ui ib ib Ul LO LO LO Ul ui ib ib ib ui ui ib ui ib ib ui ib ui ui ui LO IO LO IO IΩ LO CO CO CD CO CO CO CO CO CO CO CO CO CO CO CO CD CO CO CO C o oό oό 00 O 00 00 00 00 CO oό 00 oό 00 oό oό 00 on oό oό c oo co o co oo CO 0000 00 00 00 CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD C
CM CM CM CN CM CM CN CM CM CN CM CM CM CM CN CM CM CM CM CVI CM CM CM CM CM CN CVI CM CM sf sf st st st sf st sf sf sf sf sf sf sf f st sf st st s CD CD CD CD CD CO CD CD CO CD CD CD CD CD CD CD CD CO CO CO CD CD CD CO CD CD CO CD CD CD CD CO CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD C
© O O O O O O O O O O O O O O O O O o o o o o ooo O O O o o o o O O O O O O CO CO CO CO O CO CO CO CO O CO CO CO CO CO CO CO CO CO C o — LO Ul IO Ul U LO LO Ul Ul Ul Ul Ul U Ul U Ul Ul Ul Ul U sf sf sf st st st sf sf st st sf st sf st sf st sf sf sf st st st sf sf st sf sf st st sf sf sf st st sf st sf sf st st st st st sf sf st sf O O O O O O O O O O O O O O O O O O O
U IO IΛ U IO LO LO LO LO U IO LO IΛ IΛ IO LO LO IO UI U IO U IO IΛ LO U IΛ IO LO UI IO UI IO UI IΛ LO UI UI UI UI UI U UI UI U UI U CO CD CD sf sf st st st st st st st st sf st st st st st st sf st st st st st st st st st st sf st st st st st st st st st st st st st st st st st sf st st s^
C
cvi Cv oi oi oό C
CD CD CD CO CO CO CD CD CD CO CO CD CO CD CD CO CO CD CO CO CO CO CD CD CO CO CO CO CO CO CO CO CO CO CO S S S S S S S S S S S S S S S S S S S S S S S S S 00 00 00 C000 00 0 st sf st st st st st st st st st st st st st st sf st sf sf sf sf sf sf sf st st sf sf sf sf sf st sf sf sf sf st sf st sf st sf st st st sf sf st sf sf st sf sf st st st st st sf st st _ ) X)
H
— CD LO CO CO S sf O CD CO S CM Ol S CO sf CD CO CD O S - CM CD O LO OO CO CD CD CO S S S CO O CN — O C0 O S CD — sf S CO — sf O CD — CO sf CO O O LO CM OO CO CO CD — s CO st lO O — S CO O LO IΩ LO O IO LO CO LO CD LΩ IO — — S st O st LO sf Ul CD Ul — CD st sf Ul CO S S CD S S CO Ul Ul CO sf CD Ul OO O st CM - CM sf CD CD OO CD lO sf CO CD — sf LO O
— CO O CO CN — — O S CD Ol S Ol CD CD CD S CD CD CD CD CD CO CO CN CN CO CN CN CD O st S OO CJl — CO S IO S S — CM CM CN sf LO CO S CD CD CM O O O — — — st sf CM CM CO CO CO S CD CD CD CD CD CD CD S CD CO CD CO CO CO CD CO CO CO CO CO CO CO CO CD CO CO CO CO CO CO LO LO — — — CM CM CM CM CN CM CO CO CO CO CO CO CO CO CO CO st st st sf st sf sf sf sf sf sf st sf st st U
— CN — CO S O CO — OO OO O CO CD OI S O — CO CO — sf C CO Ul OO OO sf CN Ul CM CD sf sf LO CD Ul CD Ul CD — — OO CD IO S S CO OO CD O CD OO CD — CD UI O S OO O CO — S C
CO OO O O CO CD CD OO LΩ OO O O O O — CO Ul CO S CD Sf CD S OO CD S OO CO CD O CN CD S CO CD OO CD — sf CO — CD O sf oO CO — — CO OO CD CD CD — O — CO sf Ol — — CN CM sf CM C
CO OO CD CJl CD CD CD OO st st lO Ul Ul LO Ul Ul lO LO Ul LO CD CD O O CD CD CD O O S OO sf LO S OO — CM CO CO sf CD CD O O — CO IO LO CD CD S S S OO CD CD CD OI CJI O O - — — CM s
LO LO LO LO LO LO IO CO CO CO CO CO CO CO CD CO CO CO CO CO CD CO CO CO UI IO IO CO CO — CM CN — — — — CM CM CM CM CN CM CN CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO sf sf st sf st st s
, — ^ ^ ^ — • ^ ^ ^ ^ , — ^ ^ , — ^ ^ ^ ^ ^ rr co r: co I I — - I — I CO I I I I S I H CM I I I I — - I I CO I CO OIIII IILL I I II∑III II —
I s I I o I o s s o o , — I s — s s — o oo — oo o — s s o s s o o oo o s s s s s s s — , — o o s , — o s s s s o , — o , — o s , — , — oo s sf coo " '- '- '-
cb c oi c
O
CD CO CD CO CD CΩ CD CΩ CO CO CO O CD CO CO CD CO CD CD CO CO CD CO CO CO CO CD CD CD CO CD CO CO CO CO CO CD CD CO CD CO CO CD CO CΩ CO CD CΩ CO CD CO CΩ CD st st st st st st st st st sf sf sf sf sf st sf st sf st st st st st sf st st sf sf sf st st st st sf st st st sf sf sf st st st st st st sf st sf st sf sf sf sf st sf sf sf st sf sf sf st sf st
— CM CO S — OO sf CM CM O — CO LO OO sf CD S S O CD st cO S sf cO CD O — CD OO CM — sf CO sf — CD O sf cO OO — CM — — — — O O CN CD O — CN CO CO CN OO sf cO LO CD CN st — OI L CO CD O CO CN CD CD CD S CN O O CD CO O CD LO O UI CD O CM S S CD CO CD - CD CO CO — O CD OO OO st OO O sf st — CO — CO S - LO S CO CO OO — — CVI CN CO LΩ O CN CO CD CO O S — s CN CN CO CM CN CN CN CM CN CN CO C CN CN CO CN CN CN C CN C0 00 CM CN CM CM CN C CN CN CO C CO CN C C0 C CM CO CM CN CO CO CO W vo
O O O — O O CO O O O CO O CO CM O O - — CD CD Ij lΩ LO Ul CD CD CD CD st CD CD CD O CD sf sf S S 0000101 CD CD CO CD σ> 01 CD O CD CD CD CD CD CD CD CD CD CD CD - O 01 CD
© sf st st sf st st st st st st st sf st st st st st st CO CO CO CO O CO CM CM CM CN CN CM CM CM CO CM CN CN CN CM CO m CN CN CM C st CM CM W © Ul
H U T T T T T X I I I I I I I I I I O I I I I I I I I I I I XIII I I I
PH s X X ro — - , — o X s , — , — X X X s I X X X ϊ
, — s o s , —
, — s s
^ o o s o s oo s ^ s ^— s o s s o , — — - — - s o s oo s s s s o
^ — o 3 ^ a) — s o σ> s ^ s s , —
' '— '— '— '— ' "*- '— ^~ "I— ' ' ' o () o o o o υ o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o oo o o o o O O O OOO O o o o
CI) 0) 0) CD Φ O) 01 Cn θ) CD C Φ C O) Φ Φ Cn 01 CJ) 0) C Φ C0 O) O) Φ C^ sf sf sf st sf st sf st sf sf st st sf sf sf sf st st sf st sf st st sf st st st st sf sf sf st st st st sf st st st st sf st sf st st st st st st st st st sf st sf st st sf sf st sf sf sf st sf sf s t
_. X)
H
O CD st 00 C CCDDD ——— C CCDDD ——— O OO CCCDDD ssSftf CCCOOO OOO CCCMMM
S CN — S OI S OO CM CM CD CD CD CN S S S CM O - — — — — LO O sf CO — — CO CO CO — CN CO CO IO st O CO — sf oO CN O st O sf sf sf - — — CO sf oO st O CO — LO — CD UI S S
CO O OO O UI CD UI — sf oO S S — CO CO CN S S LO O O — CΩ CM CD CN CN O S O O — S O S sf O CO CO CM — O LO LO CO OI UI O O O — CD — CO CD sf O CO CD O Ul — CO CD CO S
CN sf CN sf S st LO OO lO CO lO LO S S CD CD CO OO CD — — — 00 — CD — — — 00 — — — CO — — — — CN CM CO CO CO CM CM CM CM CM CO CO CO CO CM CO CM CM CVI CO CVI CM CO CM CO CN CVJ CM CM C
— CD CN
— LO CD CM CM sf LO OO OO OO CD sf CD CD CD CD CD CO CO CO CO S CD O CM CD LO S — — CO CO CO CN CD CD — — O O O O O — O O O O — O O O O O O O O O O O O O O O O
— S OI CN — — — — CO CO CO CO CO O S CD CD O O O O Ul — CO CD S CD — CN CD O O O O — — CM sf sf sf sf sf st st st st st st st st st st st st st st st sf st sf st st sf st st sf s
CO — — CN CO CO CO CO CO CO CO CO st LO lO lO LO CD OO OO OO OO CD CD CD CD CD S S S OO OO OO OO — — —
I I I I CD co ui oi r: III coco sr - — — — — I I I I I — III I I I I -r
— OO CD O LO I LX II o o o DC X X H o s — s — X X — s s ro o , — s s s s s ^ s ^ s s o s o s o s r> s s o o , — s o s , — o s , — o oo o o o o s o co o o o o o ^— s o
'— '— '— '—
OO OO OO OO CO CO OD CO CO OO OO OO OO OO CO CO CO CD IO OO CO CD CO OO OO OO CO CO CO CO CO CO CO OO OO CO CO CJl CJl CD CD CJl CJl CD CD CD OT st st sf st sf st st st sf st st st st sf st sf st st st st sf sf sf st st sf st st st st st sf sf st st st st st st st st st st sf st sf sf sf sf sf st sf st st sf st sf sf st sf st st
vo
© © Ul
H — U sf X X X X X s IO X xxxxxxxr:XoXcNiιir:iioιiι IlroiLl ir: 11 I I I I I I I I I r XXXI X I I X I j" X I
CM O Ul CM X O sf CO 00 CD CD CM CD lO Ul O st l — O CD — — CD CO l CD Ol CN S S S CO I LO CD CD CN I CO CN CD S CM CO Ul CO l — 00 I: OO 100 Xsf CD sf CD CO CO 00 CD CO CN CM X O CD
PH 00 — CD CO o CD O CD 00 l O CΩ CO OO CO — CD CO O — S — — S CO CO CD CO st CD CO LO 00 CD S sf sf Ul S CO sf O — CD st CD O — Ul — S CD CD sf 00 O S S in s O CO CD CO CN CM
U CO CM CM s CM st st st f Ui σi CM CD Ul CM CD CD st CD lO Ul O CD CD OO S CD CD sf CO sf CO CD LO O O S CM — CD — CD S CD 00 st CM CD O CO OO S st CO CD CM Ul s S ( s(t) CO S Ul — OO
CM sf CM S s S 00 , — CO S CN LO CN S σi lO CO CO — O CO 00 CM O O CO CO Ul o CO 00 st LO CD CO CO CD CM S O OO CO S CO CM CD s CO O o CM CM CO — S
CO CD CD sf s st U) CO O CD S S σi OO CO O lO CM CO OO S O S CD — LO 00 CD sf LO — CO Ul S CO CD CD CD CO S O CM — S O CD Ul Ul O — CM S CO O CD LO CD sf CO st CO CO CO i— CO CD CO
CM S S CD Cl CD O , — O O CD sf Ul CM sf st oO S - CO CN CO — LO O CO LO — IO — CN — CO sf CM CM sf CO CO CO O sf CD 00 CD CO CD — 00 O S 00 sf CD — CD S s CD U) CO CD CO — CO
COLO LO LO CO U CO CO Ul CO CM Ul CN sf sf O — CD— Ol sf LO CM sf CM sf C LO CO CO O CO COsf st IO CD CM CM LO sf CO CM — sf sf sf CM sf CO CO — CO CO CM CO sf CM CM CO CN sf CD LO CO
O O O O o o O O O O O O o o o o o υ o o υ o o o o o o
Φ Φ φ φ φ φ φ Φ φ φ φ Φ Φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ
O O O O O O o o TJ TJ J TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ O O O O O O o o d d o d d d o odd d d d d d d d d d d d d
CN CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM Cvl CM CM CN CN CN CN CM CM CN CM CM CN CM — — — — CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM CN CM CM CM CM CM d CM C dM CoMdC dCM
CM CM CM CM CM CM CM CM CM CM CM CM cvi cvi vi Cvi CM Ci CN CM CM CM CM CM CM CM CM CM CM CM CM CM CM oό oό oό oό oό oό oό 00 00 CO 00 cό oό oό oό oό oό oό oό CO c 00 CO 00 CO oό oό oό 00 CO 00 00 00 00 0000 00 00 00 0000 CO CO 00 00 0000 CO CO CO CO 0000 CO CO 0000 CO 000000 00 LO LO U Ul o o sf st sf f sf sf sf st sf st sf sf sf st sf sf sf sf st sf sf sf sf sf st sf sf sf sf f sf sf sf o o o s o o O OOO o o o f o o o o st Ul Ul LO LO sf sf t sofosf Sf t o o sf sf sf sf sf osf St t sf o o o sf sf sf sf sf os sf sf sf st stost sf
— sf sf sf sf
O O O O O O O O O O O O O O O O O O O O O O O O O OOOO OOOO O S S S S sf sf sf sf st sf sf sf st sf sf sf sf sf t sf sf sf st sf sf sf sf t sf t st st sf st sf st sf sf CM CM CVI CM st sf st st st st st st st st st st sf st st st st sf st st sf st sf sf st st st sf st σιcDCDθioιcDCDCDθicDCDCDCDσ)σ)σ)σ)θioιooσ)σ)σ)σιoισιoιcDCDσιcDθocDσιoooo- — — — — — — — — — — — — — — — — — — — — — — — r-T- — — — st sf sf sf sf sf st sf st sf st st st st st sf sf st sf st st sf st st sf st sf st sf st st st sf sf lO IO IO lO IΛ LO IO LO IO Ul lΛ lO IO IΛ LO Ul Ul LO Ul Ul LO LO Λ t
st st CO — CM Ul CO sf lO - — CM CO O CM — — CM sf CD S CM O — S S sf O sf O - CD CO CD st sf CO O CM LO sf CD O CO O - — O CO UI CO CO — CD — st sf sf LO st CO st lO sf st O Ul CO O - CD S S CN O CO — — OO O OO OO CM CM CD O CO sf O CO — CO sf S CO O CD S CO CN CD OO CD O O CO LO CO CO OO CM CO — LO S sf OO CD CN — S — O O O O sf CM sf CO O st oO CD CO CO CN CM CM CO CO CO CO CO CO CM CO CN CM CM CM CO CN CO CN CO — CO CO CO CN CO CO CN CM CN CN CN CM CM CO CM CO CN CM CM CN CO CM CO CN CM CM CM CM CM CO CM CO CO CO CO CM CN CO CVI CM
CD CD CD CD sf CD CD S CD CD CD CD CD CD CD CD CD O CD CD O OO O O CD CD CD CD CD CD O O Ol CD CD O O CD O O O CN CD CD CD CD sf CO — — — — — — — — — — CO — CN CO — CM O O O
CN CM CM CM CO CN CN CM CM CM CM CN CM CM CM CM CM CO sf CM CO CN CO CO CM CN CM CM C CM CO O CM CM CM Crj CO CO CO O
— CM
I J" II I I I I I I I I I I I I I r_ 11 1 — . T T — ι
Ul st O CD — Ul sf OO lO l Ul O CO — — sf CO LO O CO CD - — CD CO CD CO O st CD CD CM CM CD - — I I CO 00 CO CO — - I I O 00 Ul CD 00 CD CO LO I CD 00 CD CM o Ul 00 X X r_ι— 00ιCD ist OO CD Ul OO sf CO CM st oO O CD CM O CN LO CM CM CM S S CM CD LO — CO O CD CO CN O S S O LO S — sf CD O Ul 00 00 U o ^ O CO Ul CM st CD CD CD sf CD CM CO o s CO S CM Ul IO CD — O N r- OO UI O CD CD — — S CN O CD sf st — CD OO CO — CN O Ul — CD — sf CN sf sf - 00 00 CD CN CO sf CD O CO CD sf CD sf CM CO — U S 00 CO CD LO CN — — O CO s CO O CM O) — O lO CN lO S O CD — OO CN sf O — S S CO CO CO O CD CO — S CO lO O CO O CD O sf CD OO UJ — 00 LO O O CD CM CM cn , — ^ CD ^ sf CO LO CD O O O O — — s O CD Ul S CO sf CD OO CD CN UI CO CD S O CO CO IO — UJ OO sf S S Ul CM Ul CM sf S — CO S OO sf - CD CO Ol — CO CO 00 CO O CO LO O — - CM
CN o CD CD CM O Ol CD sf S CM f CD Ol CM st U) s CD CO S O CD LO OO CD CO CD CD O CO CD st CD CD LO - CO — — LO O O S CO Ol OO sf LO Ul CO CD CO - CO O O CO LO — sf — CO S CD 00 Ul 00 O IO Ol — CO S sf CD CM CD Ul Ul cn CD 00 S U CD O sf — — CM sf Ul CM CVl sf CD CM - LO LO sf CD COsf CN CO sf sf lO sf CO — sf sf - OO CO CO CO lO sf sf — Ul CD CM st CO sf sf — —O C
CO sf CM CM — — CO CD Ul sf — — C st O *— CO CM U CO CD
vo o o o o o o o o o o υ o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o so CN CN CVI C CN CN CN CN CM CN CN CN CM CN CM CM CN CN CM CN CM CM CM CM CM CN CN CM CM vM CM CM CM CM CM CM CM M CVl CM CM CVl CVl CM CM CVI CM CVI CM CN CM CM CM CM CM CM CN CN Cvi cvi cvi cvi cM Cvi <M cvi (M CM CM Cvi cvi cvi cvi cvi cvi cvi cvi cvi cM Cvi cN CM W CM Ci cM cvi cM CM Cvi cM cvi cM cvi cvi cM cvi cvi cM' cvi cvi cM Cvi cvi cvi cvi cvi cM CM Cvi cvi cvi cvi oo oo ∞ co co oo co co co oo oo oo co oo oo co co oo co oo αj co co co co oo oo oo oo oo oo oo oo oo oo oo oo oo oo co oo oo oo oo oo oo oo oo oo oo co co oo oo oo oo co oo oo oo oo oo oo oo oo oo oo st st sf st st st sf sf st sf st st sf sf st st st st sf st st st st st st sf st st st st st sf st st st st st sf sf st st st st sf st st st st st st st sf st st st st st st st st sf st st sf sf st sf
© o o O O O O O O OOOO OOOO O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O o sf sf st sf sf sf sf sf st st sf sf st sf st sf sf sf sf sf sf sf sf sf sf sf st sf sf sf sf st sf sf sf sf sf sf sf sf sf sf sf f sf sf st sf sf sf st st st st sf st st sf sf sf sf sf st sf sf sf sf
CD CDθ) θi cD CD CD CDCD σ) σ) σ)σ) σ)σ) σ) σ) σ) σ)σ) θo σισ! CD CD CD σιoι σ) σ)σ) σ) σ) σ) σ) θi oι σιoι σ) θo oι oo oι σιcD σι o^ st st st st st sf st st st st st sf st st st st st st st st sf st sf sf st sf st st st sf st st sf st st st sf st st sf st st sf sf st st sf st sf st st s^
Table 4
411408.20.dec g2252026 396 814 51 411408.20.dec 3477204H1 387 804
411408.20.dec 6132793H1 392 724 51 411408.20.dec g 1357803 390 1071
411408.20.dec 2907546H1 395 782 51 41 1408.20.dec 4946208H1 388 685
411408.20.dec 785404H1 396 742 51 411408.20.dec 1576142H1 388 678
411408.20.dec 185321 H1 397 613 51 41 1408.20.dec g2619527 389 792
411408.20.dec 389185H1 396 743 51 41 1408.20.dec 743255H1 388 719
411408.20.dec 3534063H1 396 799 51 41 1408.20.dec 4958275H1 389 719
411408.20.dec 3996765H1 396 659 51 411408.20.dec 2552783H1 389 692
411408.20.dec 3053279H1 398 516 51 411408.20.dec 5098761 H1 389 710
411408.20.dec 1807609H1 399 736 51 411408.20.dec 1478849H1 389 680
411408.20.dec 2523052H1 399 737 51 411408.20.dec 2856522H1 392 721
411408.20.dec 4977531 H1 399 735 51 411408.20.dec 5687388H1 390 708
411408.20.dec 3858932H1 400 740 51 411408.20.dec 4523823H1 390 716
411408.20.dec 1845501 H1 398 658 51 411408.20.dec 3490029H1 390 750
411408.20.dec 816386H1 399 746 51 411408.20.dec 5046950H1 390 726
411408.20.dec 4198843H1 399 743 51 411408.20.dec 4586244H1 390 713
411408.20.dec 4042101 H1 398 724 51 411408.20.dec 2459258H1 390 708
41 1408.20.dec 3750944H1 400 758 51 411408.20.dec 1858361 H1 390 709
411408.20.dec 937287H1 400 738 51 411408.20.dec 4416909H1 389 717
411408.20.dec 3810490H1 400 740 51 411408.20.dec 4523905H1 389 703
411408.20.dec 3114955H1 401 756 51 411408.20.dec 2760918H1 390 762
411408.20.dec 4566634H1 401 753 51 411408.20.dec 4385055H1 390 713
411408.20.dec 3705807H1 402 773 51 411408.20.dec 4853641 H1 390 746
411408.20.dec 4066846H1 401 743 51 411408.20.dec 4819893H1 388 701
411408.20.dec 371959H1 403 562 51 411408.20.dec 3918493H1 390 726
411408.20.dec 4626333H1 403 736 51 411408.20.dec 4367949H1 391 703
411408.20.dec 4956823H1 403 742 51 411408.20.dec 4525882H1 391 703
411408.20.dec 3670548H1 406 742 51 411408.20.dec 4154002H1 391 718
411408.20.dec 3051835H1 403 768 51 411408.20.dec 5671105H1 392 590
411408.20.dec 029644H1 405 745 51 411408.20.dec 5574281 H1 391 703
411408.20.dec 1535449H1 404 654 51 411408.20.dec 4981071 H1 391 745
411408.20.dec 3592104H1 405 756 51 411408.20.dec 4895035H1 391 744
411408.20.dec 3993941 H2 405 748 51 411408.20.dec 5122182H1 392 727
411408.20.dec 5028325H1 406 750 51 411408.20.dec 4329866H1 391 700
411408.20.dec 3808513H1 419 811 51 411408.20.dec 5199638H1 392 544
411408.20.dec 3076490H1 423 753 51 411408.20.dec 2548578H1 391 687
411408.20.dec 5020492H1 423 750 51 411408.20.dec 4730925H1 391 717
411408.20.dec 2318857H1 423 764 51 411408.20.dec 3495911 H1 391 743
411408.20.dec 506742H1 424 565 51 411408.20.dec 2116906H1 391 713
411408.20.dec 3995849H1 423 768 51 41 1408.20.dec 2555366H1 391 691
411408.20.dec 4126641 H1 427 781 51 411408.20.dec 4544977H1 391 690
411408.20.dec 3930274H1 430 700 51 411408.20.dec 3373476H1 391 709
411408.20.dec 2463303H1 436 756 51 411408.20.dec 2966550H1 392 682
411408.20.dec 3137131 H1 436 781 51 411408.20.dec 2763323H1 391 692
411408.20.dec 767735H1 448 735 51 411408.20.dec 4874466H1 391 719
411408.20.dec g2056448 461 951 51 411408.20.dec 2548853H1 392 712
411408.20.dec 744424H1 468 749 51 411408.20.dec 4841824H1 391 746
411408.20.dec 508365H1 470 746 51 411408.20.dec 4953663H1 391 695
411408.20.dec 1541976H1 527 801 51 411408.20.dec 5115420H1 392 735
411408.20.dec 2404134H1 530 791 51 411408.20.dec 4845489H1 391 633
411408.20.dec 1597375T6 704 861 51 411408.20.dec 2684639H1 391 694
411408.20.dec 5700996H1 385 736 51 411408.20.dec 6376757H1 392 723
411408.20.dec 2470451 H1 385 716 51 411408.20.dec 2606634H1 392 698
41 1408.20.dec 4980959H1 385 717 51 411408.20.dec 3298129H1 391 700
411408.20.dec 2318026H1 386 713 51 411408.20.dec 4767905H1 390 736
41 1408.20.dec 4702372H1 385 705 51 41 1408.20.dec 2256874H1 391 704
41 1408.20.dec 6562862H1 392 830 51 411408.20.dec 4982262H1 391 726
411408.20.dec 2605436H1 386 704 51 411408.20.dec 3662084H1 391 742
411408.20.dec 3567254H1 386 650 51 411408.20.dec 4843594H1 392 718
411408.20.dec 4976438H1 386 712 51 411408.20.dec 3675093H1 392 746
411408.20.dec 1984353H1 386 724 51 411408.20.dec 4385088H1 391 712
41 1408.20.dec 5946551 H1 386 746 51 411408.20.dec 2548726H1 392 709
411408.20.dec 764555H1 387 703 51 411408.20.dec 3351335H1 390 713
411408.20.dec g 1955984 387 714 51 411408.20.dec 3442587H1 392 565
411408.20.dec 736413R1 388 859 51 411408.20.dec 3995356H2 392 745
411408.20.dec 4327844H1 391 700 51 411408.20.dec 4653711 H1 392 745
411408.20.dec 3617318H1 388 758 51 411408.20.dec 4977378 H1 392 731 CD sf sf CM st CM - — CD — CD — S S C sf oO CO CM CO CO sf S CD OO — CD CO O — CD CO CD sf cO CD CO — CO CO Ul O sf CN CD — S sf S CM OO OO CO CD — lO CO CM CD CO OO CM CM CN O O CD CO CD CD CN CM CN CM - CM sf CD CN O CN C Ul sf CD - CD sf CN - CO CO CO Ul O O CD CO O st O CM OO — CD — O O — CN — CO IO CM — CM lO S OO CD S O O sf CO S CN — CM CM CM CM CM S CD O CO CM CM CM CM — O S U1 CM O CM C CD sf CD S CO S S S S S S S S S CD S S S S S CO S CD S CO S S S S S S S S CD lO Ul Ul Ul Ul CD — CO S S — — — — — — — CO CO sf sf — — — — sf sf CO sf - — — vo
S — CO CO st st cO O CO CN — CN CM CM CM CM CM CO CO CM S — CM CM st cO CM CM CO st st ul CD O CM CO — CO sf CM — OO CO S sf st - — — 00 — — sf
© CD CD S S S S S CO OO OO OO OO CO CO OO CO CO CO OO OO OO CO CO OO OO OO OO OO OO CO OO OO OO CD st sf CD OO st st CD O S CM CM CM CN CM CM CM CN CM CM — CD CO CD CD CM CM CM — sf sf st L © CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO sf sf sf sf sf st — u sf st σ σ) σ σ)σ)σ) σ) - — — — 00000001 — — — — s s s Ul
H I I I X r- I I I I I I I I I XI I I I I I I I CM I 1 t CO I I I I I I CM st sf U u I r: X X sf X sf C
CN CD CO CO x S xsfiCD I O O CM LO — 00 CD CO CD — sf sf CD CO Ol sf CO st S S CO CM I I I CO 00 I O S CD CN S — S sf — LΩ U1 CO S CO CD O — f CO CD X O O sf — CD
PH CO — LO LO CD — O S U sf O CO CM sf CM O st LΩ — sf CD CM LΩ — S CO — LO S CO S S CO Ul — CD LO — O O CD sf CD S CM CD O) — 00 O U1 O S f CM Ul 00 CD CM — CD LΩ C
00 O CD CO CD CO — LO Ul LO 00 CD 00 S S CD CD — sf co ocoσi σio cD — CO CD CO ^ CM CD LΩ CD CN 00 CO CD sf — sf CD LO CM st CM O sf O S st LΩ LO S st sf CO CO LΩ O 00 CM
— LO CD CO sf CO U CD O Ol O S — CD 00 U O 0000 — CD OCoσicoco 00 CO — o 00 CD sf LO — CD — S — sf CD — IΩ CO CN — O CN U1 CM CM CO U Ul s — — CD CM S Ul f C
CD Ul O 00 Ul CD S CO LO S S CO CN LO CD f LO — s cvi iO Loo LO uiσ CD CM LO — 00 — CO CO CO 00 O LO CM CM — OI O — — OO OO — CM — CM CO CD Ul 00 CD LO LΩ CO CO CO CD
Ul — S O Ul CD CD sf — — CD S Ul LO S — — — CD CD — — CO O O OO CD O U S ^ CD CD LΩ sf 00 CD st O CO LΩ CO LΩ CN CD sf CD CD LO S — CO CN CM CM CM CD CD CD sf CM 00 CN C
CM CM — sf CM CO sf CD CM CO — CO st st — t sf CO — CO CM sf CN LO LO CO CN CO CM sf U CO S CM sf U LO CO CDCO COsf CD C0 C0CO C CM CN C CM CD CO CO COCO CD — CO CD CM CO o O O O o o o o o o O O O O O O o o o o o o o o o o o o o o o o o o υ o φ φ φ φ φ CD φ φ φ φ CD Φ Φ Φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ o υ o o o o o o o o o o υ o o o o o o o o o o o o o
TJ TJ TJ TJ J XJ XJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ J XI TJ J TJ TJ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ o o o o o d d d d o dd o o d odd d d d d d d d d d d d d d d TJ XJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ XJ XJ TJ TJ TJ TJ TJ TJ TJ TJ TJ XI TJ TJ
CM CM CM CM CM CM CM CM CM CM CM CM d d d d d
CN CM CM CM CM CM d cό CO CO ό oό 00 00 00 CO 00 CO 00 00 00 cb oό cb oό oό oό oό 00 oό oό cb oό oό oό CO oό oό 00 oό oό CO CO CO CO CO CO CO CO CD CΩ CO CD CD CΩ CD CO CO CO CO CO CD CD CD CO CD CD C o o o o o o o o oo o oo o o o o o o o o o o o o o o o o o o o o o S S s s coco co co cococo co co co cococo co co co coco co co co co c sf sf sf sf sf sf sf sf st f sf f f sf sf sf sf st sf sf sf sf sf st sf sf st sf sf sf st sf st f CD CD CD CD LO UI UI IΩ UI UI UI UI UI UI IΩ LΩ LΩ LO IΩ LΩ LΩ LΩ LΩ UI UI U U
Ul U) LΩ LΩ CD CD CO CD CD CD CD CO CO CO CO CO CD CD CO CO CD CO CD CO CD CD C CO CO IΩ LΩ LΩ LΩ LΩ LO IΩ LO LO LO IO IΩ IΩ LΩ IΩ IΩ LO IΩ LΩ IΩ LΩ LD L st st st st st st st sf sf sf st st st st st st sf st st st st st st st st sf st st st st sf sf st st sf st st st st st O O st sf sf st sf st sf st st st st st st sf sf sf st sf sf sf sf st s
— — — — — — — — — — — — — — — — — — — — — — — — — — — — — — T- — — — — — — — — — CN CN CM CM CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO C lΩ Ul lO lΩ lΩ lΩ lΩ Ul Ul lΩ lΩ lΩ lΩ IO lΩ lO LO IΛ IO lΩ LΩ LΩ LΩ LΩ lΩ IO LΩ LΩ LΩ lΩ lΩ LΩ LO LO LΩ lΩ LΩ LO lΩ t
_ω i H
— O CO CD Ol CO S sf CO LO Ul sf S CD OO Ul st CD CO CO CO CO CO O st CD CD CO st - sf CD CO S Ol — 0000 — S CD CD OO LO CO CO CM CO OO CD OO S CD OO OO CD CD OO OO CM — CO CN — CO CD
— — sf sf O CM - CO CN CO CO — CO CN OI CD LO CN — — sf O CO CM S CN CN CO O — CO LΩ LΩ LO LΩ CO LD — — O sf sf cO CD - O — Ul st CO sf st cO st st LO sf sf st Ul Ul st cO — CM — S S S U1 S LΩ U1 S S S S S S S CD CO S S CD S S S S S CD S S CO S S S S CD CD S S S CO CD 00 S S LO U1 — CO — st sf st st sf st st st sf sf st sf sf st st sf S CD S C
CVJ CM CM CM CO OM CO CN — CN CN CO CM CN CM CM CN C CM O — CN CM CN CO CM CN CN — CN CM CM CM CM CM CM CM CM CM CM CM CVI CM CM CO CO O CN lO O st sf S — Ul — sf S Ul CM — sf —
CD IJ) CD (D CD CJl CD ID σ) CJl CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD σ) CD σ> CD CD CD σ) CD CD CD CD CD CD CD CD CD CD CD OO CM CM CO CO CO st st st st Ul LO CD CD CO CD CM Ul Ul CO C
X XX
o o o o c> o o c> o o o c> o o vo o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o so oo on o o o o o o o o o oo o o o o o o o o o o o o o o o o o o
© o sf
UI U UI U UI IO IO LΩ U LΩ LΩ IΩ IΩ IΩ UI UI UI UI U UI UI UI UI U IΩ IO U UI LΩ LΩ LO LO IΩ IΩ IO IΛ UI UI U UI U U UI IO LO UI LO LΩ LΩ IΩ U IΩ IΩ UI UI UI UI
o o o o o o o o o o o o o o o o o o o o o o o o o o o o o υ o o o o o o o o o o o o o o o o o o o o o o o o o υ o υ o o o o o o o o Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ TJ TJ XJ TJ TJ TJ TJ TJ TJ TJ XJ XJ XJ TJ TJ XJ TJ TJ TJ TJ TJ TJ TJ TJ TJ XI TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ XJ XI TJ T^
CO CO CO CD CO CD CD CO CO CO CΩ CO CO CD CΩ CD CD CΩ CO CO CD CD CD CO CΩ CO CO CO CD CO CO CO CO CD CD CO CO CO CΩ CO CD CD CO CO ω co co co co co co co co co co co co co o co co co co co co co co co co co co co co co co co co co co co co co co co co co co co co co co co co co co c^
LO LΩ LΩ IΩ LΩ UI LO UI UI U UI LΩ UI UI LΩ LΩ LΩ UI LΩ LO LO U UI UI UI UI UI UI UI UI LO LO UI U U UI UI U UI LΩ LΩ U U UI UI UI LO LO U LO IO U IΩ IΩ UI vD CO CO CD CO CD CD CD CD CD CO CD CO CO CO CO CO CO CD CD CD CO CD CD CΩ CO CΩ CO CD CD CO CD CD CO CD CD CO CO CO CO CO CD CO CO ω
IO U1 LO IΛ LΩ IΩ LΩ IΩ LΩ LΩ IΩ U1 U1 U1 U1 U1 1Ω IΩ LΩ U1 LΩ LO IΩ IΩ IΩ IO IΩ U1 U1 U1 U1 U1 U1 LΩ U1 U1 LO U1 U1 U1 IΛ LΩ LΩ IΩ IΩ IO IΩ U) U) U1 U1 LΩ U1 1Ω U1 st st st sf sf sf st st st st st st st st st st st st sf st st sf st st sf st sf st st sf st st sf sf st sf st sf st st st sf sf st st sf st st st st st sf st st sf st sf sf st sf sf st st st st st s
CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO O CO CO CO CO CO CO M UI UI UI UI UI LΩ IΩ UI LO LO IΩ IΩ UI UI UI U U UI UI U UI UI UI LΩ IΩ UI IΩ UI UI UI UJ U I LΩ IΩ LΩ IΩ IΩ IΩ IΩ IΩ LO IΩ LΩ LΩ LΩ LΩ LO LΩ IΩ IΩ LΩ IΩ IΩ LΩ LΩ UI IO LΩ IΩ LΩ LΩ IΩ LΩ LΩ IΩ IΩ I st
0)
O LO O CD CO sf — O CN CO LO O CN CN CD CO — CD — CO sf — sf U Ol O OO — CM O — sf sf - CM sf —
CN LO LΩ CM st CD CD CN CD O CD S CN st S lO CD O OO sf CO lΩ CD — CD CVI CM — O CM OM CN CM CD CN CN — CN — CO O CO CN CM CM lΩ CN CN st CM CM OO OO CM CD CD O LΩ O Ul CO Ul CN CO CO O CN CD sf S CO OO CO sf O — CM O O LO CO O st — S CN CM — — CM — CM CN - CM CM CM CN CM sf CM CM CM CM CM Ul CD st CN CM CM — CM CN — CM CVI — O CM Ul — CM CD CO CD LO sr o cO O OO C — OO CD CD CD CO CO CD — — CD — — CD CD — — CM CN — — — — — — — — — — — — — — CD — — — — — CO sf CO - — — — — — — — — — — — st st sf st cO CO st st st sf st Ul C
S OO CO OO — — — CN CM CM CD CO CM CD U S — CN LO st CO CO CD S Ol Ul CD O sf sf Ul OO CD O CN LO CO CO sf sf st CD OO O CD CO CO CO OO CD CD C CO CO CO CO CO CO - — O UI UI UI UI CN CM CM O — 00 CO O1 — — — CO CD LO CD CD CD CD LO Ul CO CO CO CO CO CO CO LO LO Ul Ul CD CO CO Ul CO CO CD CD CD CD st st st Ul st LΩ LO Ul st sf st st Ul CD CO CD Ul LΩ S S ssss — — — sscococossss st sf CD CD CD CD CD CD CD ID CD Ol Ol CD CD OO OO OO OO OO OO OO CD CJl CD OO CO OO OO OO OO CO CO OO OO OO OO - — — — — — — — — — — — C
r-illcclimmiOvCiEL i ZZ I CD CD LO — I I lΩ sf l cO CN l cD CO OO CM st CVi rZ CO l CDCMcoIocoIσi — r:jrIIII IIII(D*
S st CD CN l I CD CD CN CO Ul Ul — l l OO sf I CD CO O CM CO 00 sf S CD — sf — S S CM — — CO OO CD OO l cO S LO O CVI CD CD — lΩ CM sf H X X X lΩ CD sf O l CM OO O lΩ CD S OO OO O CD CD CO CD OO S st oO σi sf CD CD CD — S 00 CO sf 01 Ul CO CM CM S CO CO CO LO S sf CD Ul CD st - CM Ol S st O CD CD CO CO S CD CO CD LΩ Ul Ul CVl OO O St sf O O LΩ OO OO CD CO Ul CM O sf st st S st CD st CM S CO CM OO CO st CD CD CO 00 sf O CO CM CD O OO sf LO CD st st cO CM CO Ul O S sf O CD C0 C0 C0 — CO sf O S CD LO Ul Ul CM sf CM CM sf CD Ul Ul LO O CD CO CM CO LO CD U U CO CO OO CD O S CO O CO O CO — S CD CD S CD CD τ- S S CΩ CD CΩ CO CN CO LO st S LΩ CM CM st S CΩ — CO σi CM OO LO OO CVJ S CO CO CO OO O sf O CO — CO UI CO — sf CD s sf Ul C OO OO OO - — CO O sf CD CO CD S O CD CO Ul Ul CM S CD O Ol OO sf CD sf CO O OO st st O O LO CO CD CD st - — O CD OI CO CM — — — sf st st CN S S OO S CO CD CM CJl st — — C LΩ — — CD UI UI — CM — CN LO — — CO CO CM CD — CD CM st — — — O sf - sf - Ul CVl — CN st CO CO CM CO CO sf - CD CO CO CO CM — CM CO CO LO — — — CD sf OO - sf st S CD CD Ul — CM O CDLΩ sf — S S CM — CD CDCM CD CD— CD CDLO — CM CO CO CO COCO CVI CD CDSf CO CD CDCO CD CO CO CO CO CD COOO CDCM CO CO CD— CD CDCD CO COCO OO OO OI CO CM CO CM CO CO CM CM CO CD COs o φ
IΌ CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO O CO CO M lΩ Ul LO LΩ Ul lΛ lΩ LΩ lΩ LΩ lΩ LO LΩ Ul Ul Ul lΛ Ul Ul Ul Ul lO lO LΩ LO lΩ Ul UJ IΛ IΛ LΩ Ul lΩ LO lΩ Ul LO Ul Ul Ul lΛ IO LO lΩ LΩ lO LO Ul Ul Ul Ul UJ Ul LO W
CD CO CO CM st Ul CO Ul sf S — S CD CO O CO CO CJl Ul st st lO CO S OO CO S — CVI CM OO CM O CO — — O OO O O OO Ul CO CD Ul CO CM CD O O CO OO OO CO CD sf CO S CD OO S CM — O CD CO st st st CO st st CM CO sf CO CO CM CO CM CD CO S CO CD CO CD S LO CD CD — CO CO CO CN CVI C vo
CO S CO CO st sf sf Ul CM CN — sf CO S O O O — sf - LΩ S — O CO sf CD CD sf st O LΩ CO CO CO CD sf S -
© sf sf st CO CO CO CO CO S O O O O O O - S sf sf - CD — sf S OO O CD S lO S S S S CO S S S O CVl CM sf st st OO CD CD O CD CD O CO st - CM CM CO S O — O O O CJl CD CO S s © — — — — — — — — CD — — — — — — CO CO S S CD CO CD CD CD σ) - CD CD CD CD CD OI CD CD CD CD O O) - — Ul lΩ LΩ sf st st LO st st Ul lO IΛ sf st sf sf sf st st st S S CO CO S Ul C Ul
H CN — — — — — CO U I I I I
CM I I sf I I I I T *- ■■- X I I I O CO I I I I I I ui I I III I ∑Z LΩ I I I 111 x r: XII I I I X X I I I I I
CM S CD CD O X I O Ul CO CM CO O 00 S — 00 CD CD CD 00 LO S CO S 00 CO 00 00 O sf I CD O — CO isf I r_ι—ιCOιCM st — CM CM I σi s co — o o — — O — — CO CD O — O s
PH CO — CM CD st CM O sf CO st sf sf S — O LΩ CD O CD CM — CO CM CO CM CO CO 00 S CO CO sf O CD sf CN S C CD CO 00 S CD sf — 00 Ul O O CM CO CM co u — s s oo 00 CN CO CN CO sf S CO 0
CM S CD CD CO CD CO S o sf sf — — sf CM CD CN CO O CM CM Ul sf sf 00 LO CN S LO CD LO CD CO LΩ CN CM 00 Ul CM O CD CD CO O CD LO sf LΩ Ul sf CD CO CM f O CM CM CD CD CO CO LΩ CD sf — O L
O 00 CD — Ul CM CO 00 st Ul Ol Sf — — CD LO CO CM CO LO CD CD sf LO 00 CO sf Ul O st CD O st CD CO CO CO LΩ CO sf CO sf CO CO CO CO CO CO — CO S CM CM 0 LO 00 00 S O Ul S O sf — Ul C
CM CM CO CD CM CD — CD o CD CD Ul S Ul CM CO CD S CO CD CD CO CD sf Ol sf CD S O CN — CD CD LO — CD CD CO CD U CO CO CN CO — CD CD CD CO CD Ul O CO CO 01 CD 00 O CN CD S sf U CM Ul CM O
S CD CD LO Ul sf CD — st 00 CO Ul S CD CO CD O sf CM CO LO CO CM CO S CM — S CO O — CO LO CO S CD LO U — sf LO 00 CO CO Sf Ol S CD CD CO f CM S U LΩ S sf O CD CM C
CN CD 00 CN CN CO CO CM — CM S sf CO CN CN st sf CM — — — CD CDCO COCN CO CN CO CO CO — CO CO CDCO sf — CO — — CO COCM CM CO — LO CO CN Ul — CD — — sf CO CM LO sf COCM — sf sf CM — CO — CO CM — C o υ o o o o o o o o o o o υ υ o o υ o o o o υ o o o o o o o o o o o o o o o o o o o o o o o o o o o υ o o o o o o o o o o o o o o o Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ TJ XI TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ XJ XI TI TJ d d iDcD (Dffl b (D 6 tDi (0! (Dcd !0(όcόuJ (D co co co co co coco co co co coco roco co co coroco coco co co co co co coco coco co co co ooco co coco co co co co coco co co co co co w
UI UI IO IO UI IΩ IΛ LO UI LO IΩ LΩ UI UI IΩ IΩ IO IΩ LΩ UI UI UI UI U UI IO IO U IO IΛ IΩ UI LΩ LΩ UI LO IO IΩ LO LO IO UI UI U U LO LO UI LΩ IΩ IO IΩ LO LΩ UI UI W CO O CO CO CO CO CO CO CO CD CΩ CΩ CO CO CO CΩ CO CO CD CO CO CO CO CO CO CO CO CO CO CO CO CO CD CO CO CO CO CΩ CO CO CO CD CO
UJ UI IO LΩ IΛ LΩ UI IΩ U LΩ IΩ LΩ LΩ UI LO IΩ UI U LO UI UI IO UI UI IO UI IΛ LΩ LΩ IO LΩ LΩ LΩ LO UI LΩ LO IΩ UI IO LΩ UI UI IΩ IΩ LΩ LΩ IO LΩ LΩ LO IΩ ^ sr sf sf sf st st sf st sf st st sf st st st st st st sf st sf sf st st sf st sf st sf sf st st st sf st st st st st st st sf sf st st st sf st sf st st st st st st sf sf st st sf st st st sf sf st s co co co co o co co co co co co co co co co co co co co coco co ococo co co co co co co co co o coco co co coco coco co coco co coco co w
LΩ IΩ U IO UI IO IΛ IΩ IΛ LΩ IO LΩ IΩ UI IΩ U LΩ IΩ U LO U IO UI IΩ IΩ UI IΩ U UI UI UI IO U IΛ UI UI LΩ IΩ IΩ LΩ UI IΩ UI LΩ U UI IΩ IO IΩ LΩ
X)
H co — s s ui — st o — — CM — s — — — st u σi co co
CN lΩ CO CO CD sf CD CO CN CM S CM CN — CO O CD S CO S S CO CO Ul OO LO O CN LO S CD st oO CM CD lO CM CM CM CM CM CO LO CD CD S LO CO Ul CM CM — CVI CVI CM OO S CD CD CO CO — CM CO CO CD O CM CO CO CM OO OO Ul CO CN CM O CM CN CM CM LO CM O st CD CO S S CM O Ul CO CM CO O CO CD CO CM - — CM CM CM CM CM CM — sf CD S O OO lΩ O CM CN CN CN CN O S S CO S CD CO — CO O U1 0 — CO CM CO CN CO CO CN — — — — — — — S Ol CO CO CO CO CO CO CO sf cO lΩ LΩ CO st oO CM CO — — — — — — — OO OO OO CD OO CD CO CO sf — — — — — LO CO CO CO CO CN CO CO CO sf st c
CN IΩ CM CM CM 0O CN CO CO CD CO
CD — CO CO CO — — — LO — Ul st OO CO CO S OO IO S 00 — 00 CN CD CN sf CD S OO OO OO OO CD CD CM CO CO — — CM CN sf — S CO CO CO — 00 00 O1 - sf L
O — — — — CN CN CN CD O CD O O O O st Ul CN CO CO CO CO OO CO CO CO CO CO — O — O — CD O O O O O O O CD CD O — CO CO st LO CO O O O O O CO CM CM CM CM CO O O O CN CO C
— — — — — — — — S 0O S OO 0O 0O CO LΩ LΩ CD CD CD CD CD CD CD CD CN CN CN — — — CD OJ OO CD CD — — — — — LΩ LΩ CO CO CD CD CD CD CD — — — — — CN — — — — — — — — — —
00 I I I I I CM σil σi o s corrco I lrllll lOlIIII T CO — O s T st CM X I I — - I I I 11 r: r: oo co o 111 co r: 1111 ∑ st CD CD CD CN CM CM — — CM S st sf CO — I CO — CO I OO — CD LΩ CD CO CN sf LO CD CN O CO CD X CO CN CO O CO — - CN CO l CM CO O LΩ CM l I CΩ — S CO LΩ LΩ CO CO l st O OO lΩ CO
CD CO CD — S CM LO — CN O CO OO CO CO sf sf st lΩ CO OO CD — sf cO S LΩ — — sf 00 CN O CM CO O CO CO Ul sf Ul , — U CO CD 00 CM O LO S CD — CD CN sf CM OO O S S sf O S O CD lΩ CO σ σi CO CO CN CO CN Sf CD — sf CM CD O O CD — LO CO OO S CN CM CO S st s S CO s st o LΩ IO sf s CO f CO Ul LO CN S CM sf Ul — CM sf O S S OO O CD CO CM CM CD L CO — 00 O1 - LO CO — CD OO CO OO sf O CO LO O CO IO S CO CVJ — OO sf S S Ul O s 00 O O sf on CM CO CO — - o — - st Ul CO CM CO S — CM CO O S S CD O CD OO O — O LO - s O CD O LΩ S — O LO — O σi CD CO CO CO CO CD — O CO LO CD O OO CO CO CD CM OO CM s O 00 00 OO 00 00 CM CN CO CD CN st S CO COCM CO CO CO— — αo s σi o cM Oi co iΩ — CO CO s CO S sf Ul CO - sf lO CN - O CN CN CO — st S S S CVl CO sf sf - — OO CO sf sf CD CO CM CM CD , — CO CD , — f CO CN CO sf sf o s S 0O — sf CM CO CM CD CO OO CD — sf sf S — CO CO s CO COsf — COsf CM CN CD COCD C0 CD C0 C0CO CO CO— CO LO CO CO CO CO COCO CN CO CO M CDCO CD CDCD CO CO CO CD O — COUl sf CO CN CM CD O) CO CO COCM CO — — CDsf CO — CM Ul CO o o o o o o o o o o o o o o o o υ o o o o o o o o o o o o o o o o o o o o υ o o o o o o o o o o o o o o o o o o o o o o o o o o o o vo T φJ TφJ TφJ TφJ XφJJ Tφ Tφ Tφj τφJ Xφi xφi Tφ TφJ Tφ Tφj Tφ Tφ Tφj Tφ -ηφ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ o so d ό( (D 6b(DbcDb 6v 6 6cD(D(θ tD 6tD Φ(Db d t
CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO O CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO W
LO UI UI LO LO LO U LO LO UI UI IO UI UI UI UI UI IΛ IO LO U UI IΩ LO LO UI LΩ UI U LO UI UI U UI UI UI U IΛ IΛ UI IΛ LO LΩ IΩ LΩ LO LO UI UI UI LO UI LO UI LΩ LΩ ^
© CO CO CO CO CO CD CO CD CO CO CO CO CO CO CO tø CO CO CO CD CO CO CO CO CO CO CO CO CO CD CO CD CO CO CO CO CO CO CO CO CO CO CO CO CO CD CO cO CO CO CO CO CO CO CO CO CO CΩ CO Ω o ID UI IΩ IΛ IΩ IO LΩ IΩ LΩ LO LO UI IΩ UI IΩ IΩ IO LO UI IΩ IΩ IΩ LΩ ID IΩ IΩ LΩ IΩ IΛ LΩ LΩ IΩ LO IΛ IΩ IΩ LO IO LO IΩ IΩ IΩ W st sf st sf st st sf st st st st st st st st st sf sf st sf st st st st st st st st st st st st st st st st st sf sf st st st st st st st st st sf st st coco co co co coco co co co co co co co co co co co co coco co co co co co co co co co crj coco co cococ n
LΩ LO LO LO U UI UI UI UI IΩ LΩ LΩ IO LO IΛ LO LΩ LΩ LΩ UI UI LΩ LΩ IΩ UI U U U UI LΩ LΩ IO IΩ IΩ UI LΩ LO IΩ UI UI LΩ LΩ UI U LO IΩ UI U U UI U UI U IΩ U IΩ
st C sf CN - CO Ul — CD — S CO CM CO — CO CO CO oo —
I I I
I I o
O cb cb
LΩ LΩ IO IO IO LO LΩ IΩ LΩ LΩ CD CO CO CO CO S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S LΩ LΩ LO LO LO U LO UI LO U U U U UI U UI UI UI U UI UI U UI U UI UI U U UI U UI IΩ LΩ UI LΩ LO IO UI U U UI U UJ U U UI UI UI UI U LΩ UJ IO UI UI UI UI LO UI UI UI st
_u
XI
H
— LO — S sf - O S st CN LO S S — CN CM O 00 — sf sf LO CD LO Ul st CM CM CD
LΩ CO — CD CD CN CN CN st S CN CD Ol CN CN CN — CO — st CD CN CN CO CO CN CD CO CD — — — OO CD CO CO — — — S O Ul CD S CO CO S CN S sf sf CN — CVI S S OI — sf cO CD CM CN CD CO CO CM S CO — CO sf CM CM CM O — CN O sf CN O CN CN O CN O sf CN CN — O CN sf — sf CM CD CM — CD S CO CO CO CN CN sf — CO CD OO CD O CO CM UI CD CD O CM O O — CO CO CO CO CM S UI LΩ CM CO CM CO CO CO Ol — — — — — — — Ol — — — — — — — CD — — — — — CD — Ol — CM CO U U UI CO CO CO UI UI S CD CO OO O S — S OO CD S O — CD — — — — — — CM sf CM CM LO CO st
Ul CO CM CM CM CM sf sf S CD CM st — CO CO CD CD CO sf CO — sf sf CO LO sf — CM — CO S CD OO CO CM CD S OO CD OO CO CO CO CO CO CO sf CO S st O — — — sf - sf
S CO CO O — — — — σi CD CO S CO CO CO CO CD LΩ S CO CO st st — — — — — CO — CM CVJ CM CM CM CO st — CM CD CO CO sf sf sf sf st sf Ul CO Ol CO CD Ul CD O - — — CO S C0 00 LΩ
CO CD CD S OO CO OO OO S S S S S S S S S S S CO OO OO CO OO OO CO OO OO CO OO — CM CM CM CM CM CO CO st sf st LO CD CD CO CD CD CD CO CO CO CD S S OO OO CD CD — — — CM — — — — sf
rr
o o o o o o o o o o o vo o o o o o o o o o o o O o o o o o o o o o o o o o o so st st st cb cb cb cb cb cb cb cb cb cb cb cb cb cb cb cb cb cb cb cb cb cb cb cb cb cb cb cb cb cb cb cb s s" s s s" s" S s" s" s" s" s" s s" s s" s" s s d d d d d d d o o o o o o o o o o o o o o o o o s s s s s s s
© s s s s s s s s s s s s s s s s s s s s o oo o o o o O
CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO st st sf st sf st sf st st st st st sf sf sf st st sf st sf st sf st st st sf s^ UI LΩ LO UI UI LΩ IΛ LΩ U LO U IΩ IΩ U UI U IΩ UI IΛ IO IJO IΛ IΛ IΩ IΛ LO UI IΩ LΩ IO IO IO UI IΩ IO LΩ IO IΩ U IΩ UI LΩ UJ UI UI LΩ UI LO IΩ W
cn cn cn cn cn cn cn cD CD cn cn cn cD cn cn cn cD cn cn cn cn cD cn cD CD cn cn cn cn cD cn cn cn cn cn cn cn cn cD cn cD cn cn cD CD Cn cn w
CD CO CO CO CO CO CD CO CO CD CO CO -I CO CO CD CO CD CO CO CO CO CO CO CO CO CO CO CD CD CO CO CO CD CO CO CO CD CD CO CD CO CD ∞ lo ro ro ro lo ro ro ro ro M M ro ro IO IO ro ro ro to ro ro ro ro ro ro ro ro ro ro IO IO lo ro ro ro ro ro ro ro ro o ro 4- 4- 4- 4- 4- 4_ 4- 4- 4*.4s.4- 4- 4- -»
CD CD CD CD CD CD CD 0101 01 Ol CD CD CD CD CD Ol CD 01 0101 CD CD CD CD 01 oi σ) σi CD CD CD CD 01 CD 01 CD CD CD CD 0101 ro ro ro ro ro ro ro ro ro ro ro ro IO CD CO CD CO CO CD CD CD CD CO CD CD
4- 4- 4- 4- 4- 4- 4- 4- 4- * 4- 4- 4- 4- 4- 4- s. 4- 4- 4- 4- 4- 4- 4- 4a. 4- 4-4- 4- 4*.4- 4- 4- 4- 4- 4- 4- 4- 4- 4- 4- sl sl sl sl sl sl sl sl sl sl sl sl sl CD CD CD CD CD σi σi 01 CD CD CD CD
01 CD CD D CD 01 CD CD CD Ol Ol CD CD 01 CD 0101 CD 01 010100 CD CD CD CD Ol CD CD CD CD σ) CD oi σi σι σi CD σi σi 00 oo CD CD CD CO CO CO CO CD CO CO CO CD CO CD CD CD CD CD CD CD CD CD J! CD CD
CO CO CO CO CO CO CO ω co CO O CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO O CO CO CO CO OJ O -■■ IO ro to ro ro ro ro ro ro ro ro ro
CO CO CO CO CO CO CO O CO CO o CO CO O CO CO CO CO CO CO O CO CO CO CO CO O CO CO CO CO CO CO CO CO CO CO CO ω ω CO CO CD CJI CD Ol O CD CD CD CD CJ) 01 CD CD CO O O co co co co co co co co o oo oocooo co co oooooo 00 00 OD O0 CD C» CO O0 CO C» CD ω CO CO CXI CO C» O0 CD CO C» CO 00 CO 00 CO co όo oo bo bo bo bo oo co co co co co co co co co o co co α. Q. α. α. α. α. n o. α. α. α. α. α. Q. α. α. α. x 0.0.0. α. α. α. x α. α. α. x a. x α. α. α. α. α. α. Q. CX α. a. ix ix ix _._._._.___._. ex α. Q. D. α. Q. α. α. !_._._- φ φ φ φ φ φ φ φ φ D CD φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ φ φ o o O O O o o o o o o o o o o o o o o o o o o o o o o o o o o o Φ Φ Φ Φ o o o o o o o o o o o o o o o o o o o o O o o o o o o co r si
CO OO oo ro ro
^ 4s.4- sl sl sl sl CD CD sl sl 4- 4i.4- 4- 4- 4- 4- 4- 4- 4- 4- 4- 4- 4s.4- 4- 4- CD ro r sl sl sl sl -_CjJcn 4s. roro ro -* ---' coco co ιo -k -' si siσi cn cn cn ro io ro io -' cn o o o 4-ω cyJCoιo ιo -i -' CD CD CD cn cn s| vj c»co c» c»CBCβ cn cn4- cn ιo 4- 4- 4-i ro ro vi oco cDcn -csco ω M^ r o) cn cn cn oo cD CD 4- 4- co ro co cD sj si c» ocθ si oo cn σ) c» si cJi cn c_ si 4-c» cn ro-«. cn 4j.4-CJi o θsi o -' 4- 4- cD4- co ro -•• -i si rocn ro ω -' Oo CO _ Ol _0000 _- O)Olvl CJ3 io cD -j- 4- cD Coo o rocDOo co co si o -i io cncn -' -' -'Ocn 4-si cn o cD CD - 4-oo l ^ -Ό UI M
4- 4- 4- sl sl sl sl CD CD sJ sl 4- 4s. ^ 4- 4- cn cn cn cn cD cn cn cn 4- 4- cn 4- cjι ro ω si si si si co si ω cD CD io ro co co co co co ιo co ro ro -^ -' 4- -' -' -i ∞ C0 C C0 CD CD CD CD Cn cn 4- C0 CO CD sl ffl sl lO θ r O -' θ r O CO sl O C» 4- sl O CD sl CD CJi rθ CJl CO IO I sl rø 4- 4- CO -' sl rø O rø CD CD CO CD CD O θ ω si rO 0J C0 sl C0 CD 4- 4- sl CD 4- C0 CC0D --'' ssll ssll OO 44-- 44-- OO CCDD 0000 Cn - 4- 4- CD CD O -k sl O CJ100 sl 4- OO sl CD IO O -' CJl sl O -' CD rO -' r 01 sl -' rθ CO C» O sl 4- rθ sl -' 4- O CD σι cD si o 4- 4- cn sj 4- oo si co cn c cnn 44ss.. ssii ccDD θθoo σσιι σσιι ccoo rroo ssii _v si cD 4- 4- cn -' si -- ro co co si cn cn oo 4- co -^ co ts>
cn cn cn cn cD CD CD cn cn cn cn cn cn cD CD cn cn cn cn cD Cπ cn cn cD CD Cn cn cn cn cn cD CD Cn cn cD CD CD CD Cn cn cD cn cn cn cD CD Cn c^
ro ro ro ro ro ro ro ro ro ro ro ro ro ro M M M M ro ro lo M M ro ro ro ro ro M M ro ro ro ro ro ro ro ro NJ to ro M M ro ro ro ro ro ro M w
0101 σi CD CD 01 CD CD CD 01 CD CD CD Ol CD 01 CD CD CD CD 01 01 σ σi σi CD CD CD 01 Ol CD 00 CD CD CD 01 Ol CD CD CD CD CD CD 0) Ol CD CD CD
4- 4- *.4- *■ 4s. 4- 4- 4- 4- 4- 4- 4- s 4- 4-4-4- 4- 4- 4- *. 4-4- 4- 4- 4- 4- 4- 4-4-4- 4- 4- 4- 4s. 4- 4- 4- 4*. 4s. 4-
CD CD σi σi oo σi CD CD CD CD CD Ol 01 CD CD CD CD σi O O) 01 σi σi σi CD O CD 01 CD CD CD CD CD 01 01 01 CD CD 01 CD CD CD 01 D D CD Ol
CO O CO CO co co CO CO CO O CO CO ω O CO CO J CO ro CO CO CO co co co CO CO CO CO O CO CO CO CO CO CO CO CO CO CO CO CO CO O O CO CO O
CO CO CO CO co co CO O CO CO CO ω CO O CO CO CO CO CO CO O CO co co co CO CO CO CO CO CO CO CO CO ω ω CO CO CO CO CO CO CO ω O CO CO CO bo bo bo bo bo bo 0000 CO 0000 bo bo bo 00 00 00 00 bo bo bo 00 00 CO 00 00 00 00 00 00 0000 CO CO 00 00 00 00 00 00 bo bo bo 00 00 00 00 CO
Q. α. ix ix 0.0.0. α. α. 0.0.0. o n a. α. ix'cxix n o α. α. ix'cxix n, α. n n n o α. x ix α. α. o n α, Q 0.0.0. n n n O, α. φ φ φ φ Φ Φ φ φ φ φ φ φ φ φ φ Φ Φ ω Φ φ Φ O o O o o o o o o o o o o o o Φ Φ Φ Φ Φ o o o CD Φ Φ Φ o Φo Φ φ ω φ φ φ Φ φ φ φ φ φ φ φ Φ o o o O o Cl o o o o o o o o o o o o o o
si 4- 4- 4- 4s. cD cn cn cn cn 4s.4- cn cn s « ωωωι)rM M M r - ωM vn
4- CD CD Cn Cn CDCO CD -l CD sl Cn<» O ωϋicn(jiw cjιoιoiMcrθ)wm «o -'MMvθω M4sr oMθ-'vn(jivθr vθcnvnw
CM - C OO CD C C CD CD CD C S S CO C — uCDj τt
oό oό oό cό o cb cb
CO CN LO st — CO S O CO IΩ O S — CN CO UO CO Ol sf O LO CM OO sf lO sf CD OO Ul Ul OO - — Ul sf Ul lO S LO CM CD st — S sf CM — S OO LO S LO CM sf CD CN CD CD CO CD CJl CD Ul S CM st CD CD sf O Ul CO CD OO O CO CO CO sf - CD O O 00 S — — — Ol CM Ul OO CO S sf cO Ul - O — S — CD OO CN CN S LΩ O O O — CM CO OO sf st LO CD CD CD O CO — CD — CO CD LΩ CO CΩ S L CO CN sf sf cO CO Ol CD CD CO CD CM st - S sf S CO CD UJ CD CD CO Ul CD CD CD CD sf st st Ul st CD ∞ S CN CM Ul Ul Ul Ul CO S O CM CM O CD CD CD CO CO CD rø S S S sf sf crj CO S S S S CN CM CM - st S S S LO Ul lΩ Ul S S S S S S S S CO CD st st st Ul Ul CD CD CD CD CD CD CO st st CD S — — sf st S sf sf sf sf sf st Ul st CO sf CO st C
O CM OO CM st cO CD O O — CM CM st S OO CO CM CM CM O CN CD CM lO CM S OO S OO CD CM st O S S σi CD lΩ O O CM CO st CO st oO CO LΩ Ol CD st ul O CO CD st O CO — CO CO CN O CD C O CO sf sf CO CO S S CD CD CD S CD S st CN CO O LO CO CO CO Ul S CM lO st sf LΩ Ui ω S CO CO CO CM CD S CD CO CD CO LΩ S CO CD S S Ul CO CO CD CM CM CD O O CN S S - O sf CD — CO S — O O — — CO CD CO CO CO CO CD — CO CD S sf LΩ Ul CM CO CD CD CN CM CN CM CN CN CM CM CO CM sf sf st OO CO CM CM CO CO CO S CD CD CD S CO st CD CD CO st st st CD CD S S S OO CD S S C S S S sf sf cO CO S S S S — CN — — — S S S lO Ul Ul Ul S S S S S S S S CD CD sf st sf sf st CD ω CD CD CD CD CN CO CO LO CN - — CO CO S st st st cO st st st sf cO CO CO CO C
— — 00 — — CO — — CD r~ τ" *"
I I oo CC S S2 — CO O l I I CM CO CM S l I S l I I I CM O S CD LL I I I Ilrlll CM — σicocoos — COlIOlcDSUlI CD LT I CD sf CD sf S lΩ — O CO CO O S sf S i CNHCO OcOoLoΩoHOO OoO lrzIi: sf cO l lΩ Ul CD st — CD CD — σi H CM S CO CM Ul CD S CM lΩ st CM CO CO IO IO CO CN CD CO IO S LO CO O CM — CM sf CD CO LO CN CD T- st st CO CM CD Ol O S sf S S O OO CD CVl LO sf CD CD CO - CM S OO CM CO CD OO s LO U UI CO OI CD CD - S CD CD S sf O OO CD Ul σ sf CO CM CD S CO S sf CN CD — OO sf O Ul CO CO O LO CO CN st CO CN CD O — LO CO LO O CO Ul Ul Ul CD st — sf sf CO sf L O O O LΩ S — CD O S OO S O CD OI O CO — CD O st OO CO lΩ O CD S S OO O CM CV1 CO — CO CN sf sf lΩ CD S S Ul sf Ol CN CD OO O S O CD CD CVJ O CM CO CO CO CO sf C CO CD CD st CO CN CO — CD CO O S st CD O S OO o cN CM σi σi σi iΩ S CD O LO CD CD LO O CO OO — sf Ul O CD Ul CM S CM - CD LΩ CO σi O OO OO Ul S st Ul CM S CD sf sf C
— S CN CD — CO CO O O sf — CO CO CD — Ul st lΩ — CO C — — CN O CO st — CM LO CD OO OO CVl sf — — CD OO OO sf LO LO CD LO CO CO S CO CO — CO sf O st cO st Ul OO S - O — CN CM LO s Ul Ul CDUl CD — COO O CO CDCD CO st CO Ol CD COCN — COsf CD CO Ul CD C01Ω COO O CO Ul CD CO CD CO— LΩ IΩ CM Ul sf CM lO CO CO CD C0 C0 CD CD C0O CO CD CD CDCM — CDCM CD CD COCΩ
O) CD CO Φ O) Φ CJ) Φ Φ Φ Φ Φ CD CO Φ C0 Φ Φ Φ ro Φ CJ) Φ O) O) Φ Φ CD Φ Φ Φ ro UI U LΩ LΩ LΩ LΩ U UI UI UI UI U IΩ IΩ IO LO U UI U UI UI UI UI UI U U LΩ UI IΩ UI U U UI UI U UI U UI UI IΩ IΛ IΩ U UI LO LO U UJ UI UI UI LΩ UJ LΩ LΩ U LO UI UI UI U)
o o o o υ o o o υ o o o o o o o o o o o o o o o o o o o
Φ cvi cvi cvi cvi cvi vi cvi cvi cvi Ci cvi cvi cvi cvi vi cvi Cvi cvi cvi cvi cvi cvi cvi cvi cvi cvi vi cb cb cb cb cb cb cb cb cb cb cb σi σi σ> oi oi oj σi oi σi σj σi σi σi σi σi oi σ> σi oi oi oi σi oi σi oi σi σi σi oi σi o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o
O — O CO — U O — CO O CD CO CD sf CO CO CN LO O CD O O CM — CN O O CD CO CN sf CD CM Ul — CD CD CO CD CO LO CD LΩ st — CO CM sf O O CO CD
CO CM CO CO OO IΩ LΩ CO CN CO CO LO O — OO LO sf — CD CO Ol OO sf OO CN OO - CO CO O sf LO S sf CO S Ul — — OO O OO st st CO CO — CO st CO OO S sf S CM S CD S CN Ul Ul CO S — CD LO I OO CM CD CD CO CD CN CO st CN — CD CN CO CVJ st CO LΩ CN CM CM CD CO OO CO O CN CN st sf CD — Ul CO OO CVl O sf CO — U O CO O — CN CO CM st CD CD O CD OO CN CD M CM CO CO st cO CO OO st Ul L CO LO S S S CO CD CO CD CΩ CO S CO CO CO CO CO CD IΩ IO IΩ S S CO CO — — — — — — sf — 0O 00 CO S — — CO CN CN CO — — — — — — — — C — — CM — CM CN CM CN CN CM CM CM CM CM C
CD lΩ CD CD CO st S CO CO CD CD — S CO LO O O — — S CD S CN S CN O S Ol S sf OO CD st CD CD S LO O CD CO LO sf CO CD Ul C
S CD CN CO sf CD sf CD CO CN CM OJ LO S - CO CO CO CO LΩ LΩ — — — CO CO — CO O CN O CO sf 00 00 sf S OO S CM CO CM OO CD sf S CO Ol CO CO S O — CO CD CD O O CM CD st CD CO CO CO CD O O — O O sf Ol CD O C CN CM CD CD CD St st OO CD CN sf O — CD S — CN S LΩ LΩ — LO — CD CN sf O O — sf Ul Ul CD S S σi O O O O O CM CM CM CM C CO sf S S S CO CD CO CO CD CD S LO LΩ CO CD CD CD sf st st S S Ul S OO CD — — S CN — C sf sf LO LO — — — LΩ — — U1 C0 CD CD — — — — — — — — — — CN CN CN CN CM CN CN CN CM C
CD •■- •■- '- — — CO '- '— τ_ CO '- — — 00 — — — — - CD '- '— '- lo — cosiio r: r: H I lO CO CN l l LO I — sf I — CO O CO l I IIU-IIIHlsIIDCIr:III_-III I I LL I I I
CO — O CO S O CN CD Ool I OIO CD O xCOx— S CO — — OO lΩ LΩ LΩ CO CO st CD C CD CN CD CO - IΩ IΩ O CD CN — CD CO — LO — CO CM sf sf l Ul — O CO CD O S CD ICN s— s CVJ 0O CO — CD — CO CN O IΩ CD U CO CN 00 O CD CD S S Ol — OM S CD CO st C CN OO σi OO CD S CO — sf CD S CO — Ul — st CO — CD S — CO LΩ S IΩ IΩ S S — CD O CO S C CO σi LΩ CD CN CO O CN sf S — LO sf — CO S O — S CD LO — O LO CO OO LΩ lΩ st CO σi S CO O CO LΩ CJl sf S CO CD CD S CO S st - CN UI O CD O CD CO CN CM OO CD CM Ul CO C O CM Ul sf CM sf - sf — CO LO CD 00 O 00 CD CN st Ol OO — OO CO UI S CO LΩ CN CD UI CD UI CO — U O CO O O U CN OO — UI OO CO CD CD CM CM OO S — st lO LO CD sf S S sf s CD CN CN sf CN O CO OO O CΩ CD 00 S O O 0000 CO S CD O O LO S CO CO CD OO O — CV1 S CM — S sf cO 0000 — LΩ C0 CD — CM OO CM CO sf CD OO OO CD CO lO LΩ OO st s CD LΩ CM lΩ sf CO S LO CN CN CM f st CD CD 0000 CO CO - CO O LO CO CD CO CO CM CM Ol — OO CO lΩ CO S S O CM CD CM CD sf st st CN sf CD CD st LO st O CN st st O CO , — CN CN CN O sf CO CN C CDsf CD CD Ol CDUl CO CDCD CD — LO CD CDLO Ul Ul CD CO CDCD CM CD CDCM CO CO COCN CD CO CO CO COsf sf CM CO CM CO LO CD CM CD CD CO LΩ LΩ CO CO CD LΩ CD sf st LO CO CD LO O s o o o o o o o c> o o c> o o o o O O O O O O O O O O O O O O O O O O O vo cn cn φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ o TJ TJ TJ TJ TJ TJ TJ T3 TJ TJ TJ XI XJ TJ TJ TJ TJ XJ XJ so oό oό cό oό oό oό oό oό oό c c oό oό oό oό oό oό oό oό ό cb cb cb cb cb cb cb cό cb cb cb cb cb cb cb cb cb cb cb cb ό vi vi vi cvi cvi cvi cvi cb cb cb cb cb cb cb cb cb cb cb cb cb cb cb cb cb cb cb cb cb cb cb cb cb c sf st sf st st sf st st sf st st st st st sf sf st st sf s oo CD OI CD CD CD CD CD CD CD CD CD CD CD CD OO OO OI CD OI O
© s s s s s s s s s s s s s s s CM CM CM CM CM CM CM CM CN CM CM CM CM CM CM CM CM CM CM C o ro ro ro ro ro o o o o o o O O O O O O O O O O O O O O O O O O O CD CD CD CD CD CD CD CD CD CD CD CD CD OI CD CD CD CD CD C
CD CD CD CD CD CD CD CD CD CD CD CD OO CD CD CD CD CD CD CD OJ CD CD O O O O O O O O O O O O O O O O O O — — — — — — — — — — — — — — — — — — — — — — — - — — lΩ LΩ LΩ lO Ul Ul lΛ IO LΩ LΩ Ul LO LO Ul Ul lO LΩ Ul LO lΩ lΩ LΩ LO CD CD CD vO CD CD CD CD CD CO CO CO CD CD CO CD CD CD CD CD CD CO CO CD CO CD ω
CM U1 LΩ CM — O sf cO S S - O S sf CM O LO S CD - S — CO CO - OO CD CO CD st st OO O LΩ CD CD Ul Ul sf O CM CD CO OO st O - sf CD lΩ S O CO LO OO CD S CO S O OO O - sf CVl lO O S O — LO O OO OO S CM CM OO CO OO st CO S — sf CM CO — CD CM CO CD S CN sf cO st oO O S — O O LO st Ul O CD CD CM CD CD CO st CN CD CO st CD S Ol LΩ O CO CD S CD CD O CD - LΩ — O O CM O — O O O O O O CD O O O — CN O O O CD O O O CD OI O CD O O O CN O CM — — — CD O — CD O CM O O O O O O O CD O O — O — O O — O O O CD CM O CM S S S S S S S S S S S CD S S S S S S S S CD S S S CD CD S CO S S S S S S S S S CD S S CO S S S CD S S S S S CD S S S S S S S S S S S CO S S S vo
IΛ JN CD CD CD CD CD CJl Cn σi CJl CD CD CD CD CD CD CD CD CD CD CD CJl st sf st st st CO st st sf st st st st sf st CD st st lΩ sf st CD IΛ st vD st CD CD l^ O O O O O O O O O O O O O O O O O O O O O — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — —
© oo oo co oo oo oo oo co co ∞ co oo co co cD cn co cD oo oo co co co co co co ω co co oo oo oo co co co co oo co co oD co oo oo oo co cD ∞ © CD CD CO CO CD CΩ CO CΩ CO CO CD CD CO CD CΩ CO CO CD CΩ CO CO CO CO CD CD CD CD CD CO CO CO CO CD CD CD CO CO CD CD CΩ CΩ CO CD CO CO CD CΩ CO CO CD CO CD CD CO CD CO Φ Ul
H — CN U i ii '-iir i iiriwilr: ': I I I CD x x - — — CD τ
X - ∑- I X X X I U-
CNιsfιCDιstιsf OxCDx —x UliOrl:ioOιCDιOOiSHOEsfxSxUlxOOιOO Srl:i-i Orl:ioOxOOrl-xCO S DC LΩ DC CD CN I O OI S I LΩ CO O CO I I OO CD S DC lO O X CO H X CO CD — CO LΩ CM C
PH S S OO CM — CM O CD CD OO — CM CD S CM sf CM - CO sf S - CD OO CM — S S CO OO CO CN OO OO LO O — sf CO CD sf OO CO sf — — CO OO CO CN CO st CD — 5. — OOΞOO1. O CD — Ul CD — CN CD CD — C Ol Ol O O sf CO S CD LO S CO S - OO CD — LΩ — CD — Ul CM O st sf CO CO - s — O O OO C CNM SS SS SS SS S CN sf — CD IΩ CO S O — S CJI O S OO OO S CO OI CN S S U sf U s S CD 00 CD CO — LΩ O CO O CN S lΩ CO st CO st OO CO CD CVJ st oO CD OO CD st CM CD CO CD CO CO U lΩ ssfr ssfr CcDo Oo Ss Oσ cD co ui — co co o s co cM CO CN LO O CO O O " -" ~ O O S CO CD — O — S CD sf — — CO CN OO O CO O O CD CD — sf O sf CD — CD Ul CO sf sf cO — CD OO S st CO CO OO sf S O lO CM CD - OO CO S LO sf sf CD CN O CD CO O O CO O CO O O CO CD OI UI O CD S O — LO CD O CO CJ σi CO CD CD S CD S CO LΩ CN O CO O - O CM OO CD LΩ sf S O S UI CD CN O O CO CD CM CD — CD LO — — CO OO — st S st LO CN CO CD OO CN CD O CD OO CD CD CO O CO — CM CO CD — CO CO CN CO CVI CO CD - — LO — — — CO CM CO CO sf S - CO CO CM CN CO CM CM S — OO CO CD CD — CN sf Ul CM CDCO — Ul S — — sf OO — CM CJI CM OO CO - CD CD — — — — CM LΩ C
O O O o o o o o o o o o o o o o υ o Φ Φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ TJ XJ XJ XJ
CM CM CM CN
01 Ol CD 01 CD CD CD CD σi σ) σι CD CD CD 01 CD CD
O O O O O O O O o o o O O O O O O r
O O O O O O O O o o o O O O O O O
CD CD CD CD CD CD CO CD CD CD CD CD CD CD
Ul U Ul l l Ul U U LΩ LO LO LO Ul U
CM CM CM CN CVI CM CM CM CM CVI CVI CM CN CM CN CM
CM CM CM CN CM CN CM CM CM CN CM CM CM CM CM CM CM CM CM CM CM CN CN CM CM CM CM CM CM C CM CM CM CM CM CM CM CM CM CM CN W CD CO CD CO CD CD CD CD CD CD CO CD CD CD CD CO CO CD CD CD CD CD CD CD CD CD CD CD CO CD CD CD CD CD CO CD CD CD CO CD CO CD CD CD CD CD CD CO CD CD CD CD ω f
OO LΩ CO LΩ S st st CO O LΩ — CN S OO O IΩ CM IΩ — sf CO S CM — CM LΩ — CN LO CO CD — O CD S LO OO CO CD CO CO st lΩ CO CD OO CD sf CD — CO CO OO O CO CN St Ul O — CO OI CO CO CM S C sf — sf O LO O Ul sf — CΩ CD CD CD CD S S CM S O CN S st CD LΩ CN OO CD O LO OO CO CN CO CM OO CD CO O — sf CD sf sf S O S CD OO S CD CD CD Ul CD CD LO CO - Ul O CD Ul O O Ol O s O CVI O O O — O O CD O O O O O O O O O O O O O CD O O CD O O O σi O O O O O CD O CM O O O O O O O — O O — O O O O O O O O CN O O O O — — CD — S S S S S S S S CD S S S S S S S S S S S S S CO S S CO S S S CD S S S S S CO S S S S S S S S S S S S S S S S S S S S S S S S S S S S CO S
CD CO OO CD (O CD CD CD CD CD OO OO CD S σ) σ) CD OO OO CD CO CO CD CO CD σ) CO CD CD CD σi OO Ol σi σ) σ) σ) CD CD CD CD σi OO O CD CD CD CD CD σ OO CD C)l CJl CD o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o — o o o o o o o o o o o o o o o o o o o o o o rø co oo oo co co oo oo co co co oo oo co oo co oo co co oo co oo oo oo co oo oo co co rø oo oo oo oo co co co co co co αo rø co oo co
CO CO CO CO CO CO CO CO CΩ CO CO CD CD CD CO CO CO CO CO CO CD CΩ CD CO CO CO CO CO CO CO CD CD CD CΩ CD CD CD CO CO CO CO CD CD CD CD CD CD CO CO CO CD
— x I I X I I I I II —I —IXI I II I I ΣZ I _- X X XI II I I I I X I I I I I I I I IIIXX I I I I X X X II
, — ^^ I s I I X o C X I — s CC s X s s ^- o o o o o o o
^ o o , — o s , — o — s o oo s s o s s s s s o o o o o o
'" '- oo σ> '- *~ *~ '— '— '— o o o o o o υ o o o o υ o o o o o o o o υ υ o o cvi cvi cvi vi cvi cvi cvi cvi cvi cvi cvi cvi cvi cvi cvi cvi vi cvi ci vi cvi cvj cvi cvi cvi CM" cvi cvi vi cvi cvi cvi cvi cvi cvi cvi cvi cvi cvi cvi cvi cvj vi cvi Cvi cvi cvi c σi oi σi oi σi oi σi oi σi σi oi oi σi oi σi oi oi oi σi σi oi σi oi oi σi σi oi oi oi σi oi oi σi oi oi σi σi oi σi oi oi oi oi σi oi σi oi σi σi oi oi σi σi σi o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o
CN CM CN CM CM C CM CN CN CM CN CN CN CN CN CN CM CN CM CM CN CN CN CN CM CM CM CM CN CM CN C CM CM CN CM CΩ CO CO CO CO CD CO CO CD CO CΩ CD CD CO CO CO CO CD CD CD CO CΩ CD CO CD vD CO CΩ CO CD CD CO CO CD CD CD CΩ CD CD CD CD CD CO ω
CD — CO O CN sf CN OO Ol S CD CO CD CN LΩ CO CO — sf — CO OO CD Ul rø CO S CO Ul CD O S CO CD st Ul CD st CD Ul CD CM CO S CD OO CD CO OO CO CN st S CO CO CO CD - CM CD O S CO — sf Ul
— CD CO Ul — CD — — CO S — — Ul — — CO CO S CO CO OO CM S — sf CO CD S — OO st st cO S S CO CO sf CO O S OO O S CO st CO S O S O O - CO O CD O O S CD S O — CO UI CD S CN O O O — CD CN O O O CN CM O O CM O O O O O O CM — CN O O CD O CN O O O O O O O O CD O — O O — — O O O O O O CN CN — O O O O — O O O — — O O O S S S S S CO S S S S S S S S S S S S S S S S S S S S CO S S S S S S S S S S CD S S S S S S S S S S S S S S S S S S S S S S S S S S S S S vo
IΛ — CM CD S S CD CO S CD sf cO CO Ul sf Ul Ul Ul Ul Ul LΩ LΩ S S CD CD σi 0i σ) CD CD CD CD CD CD σ) CD σi σi 01 CD CD CD CD CD CD CD CD CD CO CD CD CD CD CD CD OT JN sf sf — — — — sf - — st st st st st st st st st sf sf sf st st st O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O
© CO CO OO OO ∞ CO OO CO CO CO OO OO OO OO OO CO CO CO CO CD OO OO CO OO OO OO OO CO CO CO CD CO CO OO OO OO OO OO CO CO OO CO CO CO OO CO © CO CD CD CD CD CO CO CO CO CO CD CO CD CD CD CD CD CO CO CD CO CD CD CD CD CD CD CO CO CD CO CO CD CD CD CD CD CD CO CO CD CO CD CD CD CD CD CD CD CO CO CD CD CD Ul
H U I — - _- I I I I I r- 111 I I I I I I — —
I I , — o cc o I s I o
^ ^ s o o o s o s s s , — — - o o s s ^ s s o ^ s
'- '" -1- s '— '—
CVJ CM cvj CVJ cvi cvi cv σi oi σi oi oi σi o o o o o o o o o o o o o o
O — CO — S O CD CD Ul O CO OO O CD CO CO OO O st O S S sf Ul st - CO CO O sf O S LO CN CO — CD — CO 00 — CO LΩ CO IO — OO l sf CD LO CD O O Ul CO CD sf S Ul CD O CO CO O S C LΩ O CM CVI O CM IΩ CO — CD OO OO LΩ IO O CD CN S CO CN CO CO CD IO S CΩ CD IO OM O CO S CM — O CD S CO CD CN CM — CD — CD CM LO O LO — — CD CN CO CD CVJ S LO CD CO — CV1 C0 — CO CD L O O CVI CM — CM — O CN O O CD O — O O — O — O — O CD — O O O O CN CN — O — — CM O O O O O CM CM O CM O — O — O — CN — CM — O CD — O O O CM CM O CM O — S S S S S S S S S S S CD S S S S S S S S S S CD S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S CD S S S S S S S S S S
CD CD CD CD CD CD CD O CD O O — O O O O O — — — — CN CM CN CM CM Ul S CD S S S CD S S S OO OO OO OO O CM CM CM CO sf S S S CO CD S CD st CD CD S S CO S S OO O O O — — — — — — — — CN — CM CM CM CM CM CM CM CM CM CM CM CM CN CN CM CM CM CM CM C CM CM CM CM CM CM CM CM CM CN CN C CO CO CO CO CO CO C^ oo ca co oo co co oo oo co o co co αi oo oo co co oo oo co co co co co oo oo ra
CO CD CD CO CD CD CO CD CD CD CD CD CD CO CD CD CO CD CD CD CO CΩ CO CO CO CD CD CD CD CO CO CO CO CD CD CO CO CO CD CO CD CD CD CO ω
cvi cvj cvi cvi cvi cvi CVJ CM σi σi oi oi σi σi oi oi oi oi oi oi oi σi o o o o o o o o o o o o o o o o o o o o o o o o o o o o
CD Ul Ul sf sf Ul CD - CO CO CO CO sf Ul OO Ul - CM CM CD O CD S CD — CD Ol st sf — CM CM UI CO LO O S — CD CO CM O UI CD CN — O OO C CD — CD CN O LΩ O S CO CM CO ID — CO CO CD CD U
CD sf — CO Ul — st CO OO Ul lO CN OO CN CN S O LΩ S — CO — CO — O CM sf S CO CD O O sf sf CD lΩ LΩ — CO CD sf CO OO CN LO CD O CD CO O OO S CN O CN CO CD CO st S CN O C.M, C „M S . LO U _
O O CN O O Ovl O O O O O O O O O O — O O) CM 01 T- — O — O O S S CO S OI OO S CO S S OO S S CO S OO OO OO OO CD — OO CD — — CO CVI CD CD OO CD CO OO CD CD CD CM - OO C S S S S S S S S S S S S S S S S S S CD S CD S S S S S S CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CO CD CD CD CO S CD CO S S CO S CD CD CD CD CD CD CD CD CD S S CD C vo
IΛ JN — CM CM sf CM CM CM — CVl CM CO CM CM CO sf CO CO CO CO CO sf CO CO CO CO CO st cO Ul st — U CD CN UI UI UI — CO UI O CN CM CD CD CN O — sf — CM CO UI UI O O — — — OO CO CD OO — — — C
— — — — — — — — — — — — — — — — — — — — — — — — — — — CM CM CO sf Ul Ul S S S S CO OO OO CD CD CD CD CD O CN CM CM sf st st st sf Ul Ul Ul Ul Ul Ul CD CD CD S S S
© CO OO OO C» CO CO CO CO CO CO CO CO CO OO CO OO OO CO OO CO CO CO CO CO CO CO OO IO U1 U} U) LO U1 U1 U1 U1 U1 U1 U1 U1 U1 U1 U1 IO IΛ CO CD CO CD O © CO CO CO CO CO CO CO CO CD CO CO CO CD CO CO CO CO CD CO CO CO CD CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO O CO CO CO CO CΩ CO CO CO CO CΩ CO CO CO CO CD CΩ CO CO CO Ul
H U I ∑z — I I — I I X X X sr I 11 r_ I — I — —
PH s cc s s I o — , — - I I , — o s o X H X U s o s , — s oo s , — s s s o s o , —
, — , — s s o s o s , — o o s s s , — s s o s o o s oo o s o o s s , — , — —
*~ '" '- '- τ_ '- τ_ '" '- '- '- o o o o o o o o o o o o o o o o o o o o o o o o o o o
Φ
Cvi cvi cvj cvi CM Cvi cvi cvj Cvi cvi cvi Cvi cvi Ci cvi cvi cvi cvi cvi Cvi cvi cvi vi cvi cvi i cvi cvi cvj Cvi cvi cvi cvi cvj cvj cvi CVJ cvi cvj cvi cvi cvj cvj Cvi Cvi Cvi cvi cvi cvi cvi c σi σi σi oi σi σi σi oi σi oi oi oi σi oi oi σi oi oi oi σi σi σi oi oi oi oi σi σi oi oi oi oi σi σi σi σi oi oi σi σi σi oi σi σi σi σi σi σi oi oi σi oi σi σi σi σi o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o
CM CN CN CM CN CM CN CM CN CN CN CM CN CM CM CN CN CM CN CM CM CM CN CM CN CN CM CM CN CM CN CN CN CM CM CM CN CN CM CM CO CD CO CO CD CO CD CO CD CD CO CO CO CD CD CO CO CO CO CO CO CO CO CO CO CΩ CO CO CD CD CD CO CO CD CO CΩ CD CO CO CO CO CO CO CO CO CO CO CO vD CO CO CO CO CO CO CO CO CO CO t
X)
H
LΩ — CD Ul — CD O CN O — CO CO CO CD CN — sf S CO — S — CVl O st CM Ul — CO CO OO O Ul — CO st CM CD CO S CN CN CO lΩ CM CD lΩ S O CO LΩ — CD CO sf O OO S LO — S sf C COO SS SS CO s CD S LO Ul CJl CD CD CD CD O CD O st st sf CD S C Ul S st CD CO CO CD st O O CO CO CO CO CO Ul S O CD OO st O S Ul st CD CN OO OO CD O S S OO S S CN S S - — CO CN OO S — CO " 00 C O O O O O O CD O O O O OO CD CD O CD O O O O O O O — O O — — O O O O O O O — O CD O — O O O CD — O O O — O 01 — O O O O O O CM O — O — CM O o S S S S S S CD S S S S CO CD CD S CΩ S S S S S S S S S S S S S S S S S S S S S CO S S S S S CD S S S S S S CO S S S S S S S S S S S S S S s
CD O) 0) 0) CJ) CD CO CD OI CD CD CD CD CJ) CJI T— o σ) cD θ o σ) σ) σ) τ- σ) θ σι o o o σι σ) θ σ) θ θ o σι σ) σ) — o σι σ) σι σ) θ θ τ- st τ- τ- τ- τ- τ- τ- τ- cM θ τ— θ τ— - — T- T ooooooooooooooo — — OO — — OOO — O — O — — — OO — O — — — OOO — — OOOO — — — — — — — — — — — — — — — — — — — cooooorøoococooococorøcoαDcococo∞∞cocorøoocococococococo
CΩ CO CO CO CO CO CO CO CO CO CO CO CD CO CO CO CO CO CO CO CD CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CΩ CO CΩ CO CD CD CO CO CO CO CO CD CO CO
I
o o o o o o υ o o o o o o o o o o o o o o o o o o o υ o o o o o o o o vo o so cvi cvi cvi cvi cvi cvi vj cvi cvi cvi cvi cvi Cvi cvi cvi cvi vi cvi cvi cvi cvi cvi cvi cvi Ci cvi Cvi cvi cvi cvi cvi CM cvi cvi cvi Cvi Cvi cvi cvi cvi vi cvi cvi cvi CM vi cvi cvi cvi cvi cvi cvi cvi cvi cv σi oi oi σi oi σi oi oi oi oi σi σi oi oi oi oi oi oi oi σi oi oi oi oi oi oi oi σi oi oi oi oi σi oi σi oi oi σi oi σi σi oi oi oi oi oi oi oi σi oi σi oi oi oi oi oi oi oi o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o
© o
CM CN CM CM CM CN CM CN CM CN CN CM CN CN CN CM CM CM CN CN CU CN CN CM CM CN CM CN CN CM CN CN CM CN CM CM CN CO CO CO CO CD CD CD CO CD CD CD CD CO CD CD CD CO CD CO CO CO CO CO CD CO CO CD CD CΩ CD CD CΩ CD CO CD CD CD CD CD CO CD CD ω
Table 4
256009.2.dec 870733H1 6672 6917 62 256009.2.dec 1301095H1 6800 7056
256009.2.dec 990641 H1 6674 6982 62 256009.2.dec 3389837H1 6800 7061
256009.2.dec 947048H1 6677 6832 62 256009.2.dec 4229811H1 6800 7084
256009.2.dec 206785H1 6677 6908 62 256009.2.dec g1365272 6801 7224
256009.2.dec g1506961 6694 6872 62 256009.2.dec 933491T1 6800 7177
256009.2.dec 5006831 H1 6694 6807 62 256009.2.dec 1955431 H1 6800 7075
256009.2.dec 3699713H1 6694 6856 62 256009.2.dec 2265155H1 6801 7070
256009.2.dec 4824777H1 6694 6925 62 256009.2.dec 3352558H1 6801 7093
256009.2.dec 4407873H1 6694 6956 62 256009.2.dec 4242674H1 6800 6987
256009.2.dec 3706839H1 6696 6968 62 256009.2.dec 555446H1 6802 7043
256009.2.dec 663677H1 6696 6935 62 256009.2.dec 933491 H1 6800 7072
256009.2.dec 1456470H1 6698 6972 62 256009.2.dec 2972455H2 6801 7098
256009.2.dec 968683H1 6703 6976 62 256009.2.dec 3726151 H1 6802 6923
256009.2.dec 6221267H1 6704 6979 62 256009.2.dec 3369242H1 6802 7046
256009.2.dec 2099129H1 6706 6960 62 256009.2.dec 4880939H1 6802 7080
256009.2.dec 2429628H1 6706 6939 62 256009.2.dec 953686R1 6802 7215
256009.2.dec 895220R1 6707 7113 62 256009.2.dec 1445580H1 6802 7043
256009.2.dec 895220H1 6707 6926 62 256009.2.dec 2530970H1 6802 7032
256009.2.dec 2971992H1 6709 7001 62 256009.2.dec 2639476H1 6802 7050
256009.2.dec 2354602H1 6713 6935 62 256009.2.dec 953686H1 6802 7045
256009.2.dec 3441955H1 6713 6959 62 256009.2.dec 4507544H1 6802 7079
256009.2.dec 3704091 H1 6713 6980 62 256009.2.dec 3726182H1 6802 7055
256009.2.dec 3705591 H1 6714 6991 62 256009.2.dec g1390836 6803 7209
256009.2.dec 3481169H1 6714 6854 62 256009.2.dec 4815838H1 6802 7057
256009.2.dec 2647507H1 6717 6971 62 256009.2.dec 2508350H1 6803 7057
256009.2.dec 2260986H1 6722 6943 62 256009.2.dec 2061029H1 6802 7061
256009.2.dec 2366005H1 6723 6956 62 256009.2.dec 2910492H1 6803 7052
256009.2.dec 344692H1 6730 6929 62 256009.2.dec 2061029R6 6802 7124
256009.2.dec 4629622H1 6730 6990 62 256009.2.dec g5112634 6803 7218
256009.2.dec 4625565T6 6751 7201 62 256009.2.dec 3282530H1 6804 7046
256009.2.dec 4383604H1 6755 7001 62 256009.2.dec 4879137H1 5767 6051
256009.2.dec 385999H1 6763 7029 62 256009.2.dec 2579142H1 5789 6048
256009.2.dec 2913295H2 6764 7035 62 256009.2.dec 3286303H2 5797 6063
256009.2.dec 478883H1 6764 7049 62 256009.2.dec 3405455H1 5802 6041
256009.2.dec 3287375H1 6767 7010 62 256009.2.dec 6521704H1 5808 5885
256009.2.dec 6311868H1 6767 7230 62 256009.2.dec 1986283H1 5822 6032
256009.2.dec 2903120H1 6765 7068 62 256009.2.dec 2911956H1 5829 6113
256009.2.dec 6484690H1 6769 7215 62 256009.2.dec 1403913H1 5833 6084
256009.2.dec 1727418H1 6779 6987 62 256009.2.dec 5437246H1 5842 6082
256009.2.dec 4654772H1 6780 7027 62 256009.2.dec 6492553H1 5851 6377
256009.2.dec 666300H1 6787 6997 62 256009.2.dec 3025734H1 5852 6097
256009.2.dec g921826 6789 7192 62 256009.2.dec 4456910H1 5865 6130
256009.2.dec 4351905H1 6788 6948 62 256009.2.dec 5093151 H1 5870 6132
256009.2.dec 1596705H1 6789 7011 62 256009.2.dec 3423758H1 5871 6126
256009.2.dec 2998188H1 6789 7046 62 256009.2.dec 3761955H1 5880 6181
256009.2.dec 2998354H1 6789 7047 62 256009.2.dec 5301333H1 5882 6083
256009.2.dec 5450729H1 6789 7059 62 256009.2.dec 6380485H1 5882 6153
256009.2.dec 1406482H1 6789 7024 62 256009.2.dec 284271OH1 5882 6164
256009.2.dec g864197 6793 7137 62 256009.2.dec 4138977H1 5885 6162
256009.2.dec 1880440H1 6794 7039 62 256009.2.dec 4353518H1 5887 6065
256009.2.dec g995262 6795 7200 62 256009.2.dec 5048915H1 5897 6153
256009.2.dec g1847140 6795 7184 62 256009.2.dec 4874435H1 5903 6170
256009.2.dec 5328912H1 6797 7039 62 256009.2.dec 2552601 H1 5913 6166
256009.2.dec g2005208 6798 7103 62 256009.2.dec 2970306H2 5921 6239
256009.2.dec 862304T1 6797 7167 62 256009.2.dec 2095611H1 5922 6188
256009.2.dec 4382949H1 6798 7052 62 256009.2.dec 2253545H1 5921 6158
256009.2.dec 1209525T1 6797 7174 62 256009.2.dec 3404668H1 5922 6161
256009.2.dec 545331 H1 6796 7032 62 256009.2.dec 3886564H1 5927 6163
256009.2.dec 4070318H1 6801 6924 62 256009.2.dec 5550359H1 5929 6178
256009.2.dec 567575H1 6799 7058 62 256009.2.dec 366241 H1 5930 6121
256009.2.dec 2180691 H1 6799 7054 62 256009.2.dec 5173270H1 5948 6224
256009.2.dec 1559402H1 6799 7018 62 256009.2.dec 3281852H1 5965 6211
256009.2.dec 3487562H1 6799 7084 62 256009.2.dec 3678857H1 5980 6209
256009.2.dec g5397606 6799 7217 62 256009.2.dec 3321667H1 5992 6276
256009.2.dec 5182883H1 6800 6959 62 256009.2.dec 2095264H1 5996 6272
256009.2.dec 5478393H1 6799 7079 62 256009.2.dec 4361881 H1 5999 6249
256009.2.dec 933491 R1 6800 7215 62 256009.2.dec 4357478H1 6014 6094 o υ o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o υ o o o o o o o o o o
Φ cvi Cvj cvi cvi cvi cvi CVJ CVJ ci cvj vj cvj cM- cvi cvi cvi cvi cvi cvi cvi cvi cvi cvi cvi cvi CM CVJ CM cvj cvj cvj cvi cvi cvi Cvi cvi cvi cvj cvj cvj cvi cvj cvi cvi vi cvi Cvi C oi oi oi σi σi oi σi σi oi σi oi σi σi σi oi oi oi oi cD oi σi σi σ> σi oi oi CD" σi oi oi σi oi oi oi σi oi oi σi σi σi CD" oi oi σi σi σi σi oi oi oi σi σi σi σi oi σi oi o o o o o o o o o o o o o o O o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o
CN <M CN <M CN CN CM CN CM CN CM CN CM CM CM CN CN CN CN CM CM CN OM CN CN CM CM CN CN CN 0M CM CN CN CM <N CN CN CN CN W CD CO CO CO CO CO CD CO CD CO CO CO CO CO CO CD CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CD CO CO CO CO CO CO CO CO CO CO CO CO CO CD CO CΩ CO CD CD CO tø t _
X)
O CD — S S CM CD CN CD Ul O O CO CD st CO st st CD st st st sf CO S st O Ul CM CD S st — CD — sf sf CD O O Ul LO sf S CO CD sf CD — O O CO LΩ CO CN st — OO CO CO CD sf cO — O OO I — O S CD sf CO — CN O O LO CD — CO CO S S S CVJ CO CO OO CO S CO CO O sf — CD CD OO Ol O OO OO OO O CN S O st OO CO CD LO CO CD CD CD S Ul CO S CM CVl S CO Ul CO CO CO CO CD - CO C CO CN CO CM CM CO CO CO CD CM LΩ CO CO CO CO CO CD CO CO <r> CO CO CO st CO S st CO sf st m st Ul sf st M Ul CO M U) U) Ul CO CD CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CD CO CO CO CO CO CO CO CD CO CO CO Φ
O U1 C000 C0 C0 C0 LO O — st CD LΩ OO CN CM CO CO st CD CD OO CD CO OO CD OO LO Ul S CO CD CO st S — S CM sf Ul CM CO sf Ul OO OO Ul — O CO CO sf S CO O O — CD CO — — — — — — —
CN CM CO CO CO UI U CD S CD — — CN CO IΩ LO IΩ LO LO LO IO LO LO CD CD CΩ S CO CD CJI O O O — — CO CO st LO LO CO CO CO OO CO CO CD O — — — — — — CM CM CN CO O — — — — — — —
O O O O O O O O O O — — — — — — — — — — — — — — — — — — — — (M CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM CO CO CO CO CO CO CO CO CO CO CO sf st sf sf st sf st sf s
CO CΩ CO CD CO CD CO CO CO CO CO CO CD CO CO CO CD CO CO CO CO CO CO CO CD CO CO CO CD CO CO CO CØ CO CO CO CO CO CO CO CO CO CO CO CØ
Hr: u- I _- II I I I I I ΣZ IX I 1 I I I I I I I I r:i sf i I I I I ∑z I ∑ I ΣZ I I I I I I I I X X - X X X xi x ∑
CD S I CD CD I sf CM co co I 00 CD sf X X rririi II I
I CD I O CM 00 O st st Ul LO LO CO O CD X sf CM CO CM sf CO CN I CD I CD I LΩ CO CM — CM — sf st ui co X CD CM CM CO sf CO
LO S CN sf s co s sf st Ul U CD CD — CD CO U U 0000 en ^- Ul CM — CO CO S — 00 — CM CO s co σi σ) sf S CO CD 00 CO sf sf O LΩ O O Sf LΩ st O st CO 00 CD CD sf S CO sf CM S S S 00 C
CD O Ul CO CO CM CO CN sf — 00 CD CO — CO CO o o o O — Ul Ul CO CM CM CO CD CD st sf LO O O st sf sf O S Ul CO CD CN O LO CM CO O — CO O O CD 00 O sf sf OO sf CD st S CD CN L
00 O CM — — CN CO f Ul CD 01 — CN sf — CO CO s s CD — CO Ul O CO CO Ul CD sf Ul CO sf — st CO CO CO CM O f CD — U O CD S 00 U — sf CO 00 CM CN CD S Ul 00 Ul CD S CD U CN C
CO CD CM O O sf CD S O sf CO — s S CM CO CM s S S CM CD st CD O CM S 00 Ul CO CM 00 CO 0 CN CD OO U S O — CD — Ul CM CN CO sf — S — CO CO U O sf st CN U st CD Ul O
CD S LO CO CD — sf LΩ Ul CM Ul sf CO CM — CO S 00 — — 0000 sf CO CO CO 00 CO — S CO — CD st — CD — S t CD 00 O CD CO CO CN O 00 CO CM sf sf 00 Ol CD CM CM sf o O CO — C
CO CN LΩ — — CN Ul sf U CD sf CO CO sf CM Ul — — CO sf S sf CM CO sf sf CM CO st LΩ CO sf LΩ 01 CO CDCO sf CO CD CM S sf CD sf S CO CO — st Ul U U sf — CM st CM — CO Ul CM CN s o o o o o o o o o o o o o o o c> o o o o o o o o o vo φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ CD φ o TJ TJ TJ TJ TJ TJ TJ TJ TJ XI J so
CD CD CD CD CD CD CD CD σi σi σi CD CD CD O) σ> O O O O O O O O ooo O O O O o o ooo O O O O O O O O O ooo o o O O o o ooo O
© CD CD CO CD CD CD CO CO CO CD CD CD CD CO CD CD C o LΩ LΩ LΩ U Ul Ul Ul U LO Ul Ul U Ul Ul Ul Ul U CN CM CM CM CM CM CM CM CN CM CM CM CM CM CM
o cvi C cvj cvi cvj cvi Cvj Cvi cvi cvi cvi vj cvi cvi Ci C oi o o o o oo
CM CN CN CN CN CN CN CN CN CN CN CM CM CM CN CM CM CM CN CM CM CM CM CN CM CN CN CN I CN CM CM CN CN CM CM CM W CO CO CO CD CO CO CO CO CO CO CO CO CO CO CO CO CO CD CO CO CO CO CO CO CO CD CO CO CO CO CO CO CO CO CO CO CO CO CO CO CD CD CO CO CO CO CO CO CO CO CO CO CO CO CD tø
X) ca
H
LΩ LΩ st CM CO S st CN CD LΩ OO CO O CO CO — LO O CO OO OO CO CO CM OO Ul S CO S CO O CO CM O st Ul st S CM Ul OO st D st CM CM CD Ul S O O S CM Ul st S - CO IO O O CO — S CO O CD st CD CN — sf sf sf — CM Ul sf CD sf sf sf Ul Ul CO CD - (M O S S OO O OO CD CN CO σi st - O — CD S OO st CM O O OO CO CO S CO C CO CO S OO S O OO CO OO S O CO O O O CD O CD — CD — O — sf CM CN CM CM CM CM CO CO CO CM CO sf st CO sf st st st cO Ul lΩ st sf LΩ lΩ CD CD CD CD CD CD CD CD C sf sf sf sf sf ssrf sstf st sr st sf st st sf -l st st sf st Ul st Ul LΩ Ul sf Ul st Ul st Ul Ul Ul Ul Ul Ul Ul Ul Ul Ul lΩ lΩ lΩ LΩ LΩ LO LO Ul Ul Ul Ul lΩ Ul Ul Ul LΩ LΩ Ul LΩ lΩ lΩ LΩ LΩ lΩ W
CO. COj CD. U -, -Ul. O LΩ LΩ CO — sf CD OO — — sf CM sf - CO O CD — CD CO OO — S CO CO O CD LO Ol Ol sf CJl CN st S - OO S OO CN OO st — CN CO CN CD CD Ul Ul O CN O CD sf CO — CO CO CD CM C S vD CD O CN CO LΩ S OO O CO st CD st sf S O CO st CO O O CN S S S CD CD S OO CD CD sf LO Ul S OO CVI CO Ul CO OO CD O — CO Ul CO OO O CM st lΩ CD CD CD O CO Ul CO S OO CO O O — C — — — CN CVJ CM CO CO CO sf st st st Ul Ul Ul CD CD CD CD S S S S S S S CO OO OO OO OO CD CD CD CD CD O O O O O O — — — — — — CN CM CM CM CM CVI CM CO CO CO CO CO CO CO sf sf st s st st st st st st st st st st st st st st sf st sf st st st st sf st st st st st st sf sf sf sf sf st st st st Ul Ul Ul Ul Ul lO IO IΛ Ul LO Ul Ul Ul LO Ul Ul Ul Ul lΛ LO lO LΩ W
— CM — 00 —
I I I X I I I — — CD
I I I X r. i X I I LL X I I I I I I
CO S O S st O I— Isf IUl I — Isf ICM Isf Isf I— Ul ICO II I I I I I I — I I II I I IXII X I I I I I X X X
CO CO 00 — S CD CD s — CVJ CM Ul Isf s — 00 S CD CD CD O C
— Ul co o s oo s I
— O — — sf Ul — 00 LΩ S CO CN 00 O CN CD 00 CO LO CO CD CO sf CD CD CM s O 00 0000 LO s s s s CVJ CVl Ul CO CD st st s CD U st S st CD CO sf — CM S CO O 00 CD CD CO S O CD — CO 00 O — CO — LΩ o CD CN CM CM sf s lΩ st CO O CO CN Ul C
— LΩ S st CM CD O CO U CN CO CN CD — CO sf CM 00 CD O S sf CD — CM O CD CM CO O O U) Ul oo o CD CD S Ol st lΩ OI C sf CD O st U CO st sf 00 CO — CO CN 0100 CN C 00 O — — OO S LO CD CM CM CM C sf sf O CM o sf CD CM Ul st st st C sf O sf CD CO CO CD CD CM st — — Ul sf sf — — S — CD U CM O CO — 00 sf LΩ LO S S 00 s o s s CO CO CN st CO OO CD C
CO Ul CD Ul CO CD Ul CM CO CO CD CN CO CO CD CN CM CO Ul CD— sf CD sf CM sf CO LΩ CO CM CD — CD CO CO CN CO LO CN CO sf CM CM sf s
O O O o o o o o o o o o o o o o o o o o o o o o O O O O O o o o o o o o o o O O O O O o o o o o o o o o o o φ φ φ φ φ vo φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ Φ φ φ φ φ φ φ φ φ φ φ φ φ oφ oφ o TJ TJ TJ TJ TJ XJ XJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ XJ TJ TJ XI XJ XJ XJ TJ TJ TJ TJ TJ TJ TJ TJ TJ XJ XI TJ TJ TJ J TJ so CM CM CM CM CM CM CN CM CM CN CN CN CN CM CM CN CM CN CM CM CM CN CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM CΪ CN i C- CM CN CN CN CN CN CN CN CN CM CM CM CM CN CM
CD D σ) CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD 01 CD CD CD CD CD CD O CD CD CD Ol CD CD CD CD CD CD CD O) CD CD 01 CD CD CD 0000 CD CD CD CD CD σ oOoO O O O O O O O O O O O O O OOO O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O
O O O O O O O O O O O
CD COo O O O O CO oo)o oo O O O O O OOO O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O o σi oo
CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CO CO CD CD CD CD CD CO CO CD CD CO CO CD CD CD CD CD CD CD CD CD CD CO CD CO CO CD CD CO CD CO CD CD CDoo
© CD CD CD CD C Ul Ul l LO Ul LΩ U LO LO LO LO LO LΩ LΩ Ul U U Ul Ul Ul Ul Ul CO Ul LO Ul Ul U) Ul U Ul Ul LO LΩ LΩ LΩ LO U LΩ Ul LΩ LO Ul Ul U Ul Ul Ul U Ul Ul U Ul Ul LO LO Ul LO L o CM CM CM CM CVJ CM CN CN CM CN CM CM CM CM CM CM CM CM CM CN CM CN CM CM CM CM CM CM CM CN CM CN CM CN CN CM CM CM CM CM CN CM CM CM CM CM CM CM CM CM CM CN CN CM CM CM CM CM C
CM CM CN CN CN CN CM CN CN CM IM CM CN CN CN CN CN CM CM CM CN CM CN CM CN CM CN CN CN CN CN CM CN CM CN CM CN CM CM W CO CO CO CD CO CO CO CO CO CO CO CO CO CO CO CO CO CO CΩ CO CO CO CO CO CO CO CO CO CO CD CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CΩ CO CΩ Φ
Ul Ul σi CO CD O st lΩ Ul LΩ st LΩ O CM LΩ LΩ Ul CO Ul sf lΩ CD S O — O — CD O CM CN CD O st CO CO sf sf O sf O CO O CD CD LO LΩ LΩ S O - CM LO sf CD LΩ
— — — — — C — — — — CD CO CN — — — — — — O — 00 — O Ol sf CD Ul sf sf CD - — O LO — OO S S sf sf CD CO CO CO - — OO sf CM Ul O CO OO O O — — — O CM — O — — CD —
CN CM CM CM CM CM CM C CM CM — CM CM CM CM CM CM CN CM CN CM CM CM S CO CO st CD CO CD CD CD CM CD CM O CM CM CO CO CN st st S CD S S S CO O ∞ S S S S S S S S S S S S S S S S S S S S S S S sf S CD OO CD S S OO OO CD OO CM - — — — — — — — — — — — — — CM — CM CM CM CM CM S S S S S S S S S S S vo
IΛ CM CO sf Ul Ul Ul lO CN CO S CD CM CD CD CD CD CD S O O O CM CM CN OO O S CD O UI CD S CO CD CO O — CM OO LO CO O CM CO S sf CD CD O O O CM st JN S S S S S S S OO CO OO CO OI OO CD CD CD CD CD O O O O O CM UI CD CO U — O — CM S O CM CD — — CN sf CN sf S CM Ul — — CD CM CO — S — CO LΩ O O O — — — CN CN CM CVI CN
© CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CJI CD CD O O O O O CD CN CO CO OO sf O O sf sf Ul CD O O - — — CM CM CO sf st Ul Ul Ul CO CD OO OO CD O O O O O O O O O O O O O © CO CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CO CD S S S S S — CO sf sf st st Ul CD CD CD CD CD S CD - — — — — — — — — — — — — — — — — C CM S S S S S S S S S S S Ul
H — — — — 00 00 U x ir: I I I 0)0)IIIr_I-III_II X sf X — I CO I LL I I I I I I I I I I I I I o X X X X I CO I I I I CO I I I Ul 2- I Z I I
SrI: σCD I LΩ rI: 00 LO CO CD — S — I — CO l S S — S CD OO 00 O C0 O 00 CN C0 LΩ Ul — Ol CO sf 00 00 Ul CO sf Ul sf st st st cD - co σi CD LO 00 — CD CD CO CO CO O CM CM I CD CC 00 CD sf S CO CO S CD Sf U — lΩ O CO CO CD CO CN CO CN OO S OO O OO CO — O O S OO CN CN CM O S S CD sf — 00 O CO — CO sf O — O CM CN — CO O CO CN — sf Ul — — 00 O CM 00 CO 00 Ul O
CD CD — CD Ul sf S 00 S U1 S — Ul Ul Ul OO Ul CO st lΩ st OO CO CN CO CN st CO O CO O O CD CD — S LO CD LO CD CN 00 Ol sf - O Ul sf S sf CD Ol st — U CD 00 S O CD S CO S CO S CM CM
CD — — — CD sf 00 f sf CO CD OO S CO S CM S sf - S — sf CD CO CD CO CO S CO IΩ CN CM O CD LΩ CM O S 00 — CN — UI CM S CM OO CN CO S O — S O O CN U CD sf CO CD CVJ — Ul — O CD
CO CO CM S — — — CN OO CD — CD CD OO CD CN CD CD Ul st Ul Ul Ul 00 — sf CM OO CO O S S CD CM CO sf CO CO CD CO sf sf C0 CD CO CO CM CD C0 CM — CD CD sf CD CD CD sf CO O CD O CD 00 CD — Ol
CO LΩ CM S LO S sf O — LO CN OI CM CD CM O CO S CM O CM O O st sf CO CM st st S O O O CD Ul CO sf CD 00 CO O CO CN sf CD CN OO O CD CO LO O U) CD O sf — — CN — 00 — 00 CD 00 sf CD
Ul 00 Olsf CO LO sf CN CO COCO — — CD — CM CD sf — CM — LΩ LΩ CO CDLO CDCD CDsf CO CO Ul CM COUl CO LO U) Ul Ul U sf CD Ul Oi σiLO LΩ lΩ U LΩ CDUl LO CO sf COCM CO — COS — S CO Ul o o o o o o o o o o o o o o o o υ o o o o o o o o o o o υ o o υ o o o o o o O O O O O o o o o o υ o o o o o υ o o o O O O φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ Φ Φ φ φ φ φ φ φ φ φ φ φ Φ Φ Φ Φ φ φ φ φ φ φ Φ Φ Φ Φ Φ φ φ φ φ Φ Φ φ φ φ φ Φ Φ φ φ φ φ φ
TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ J J TJ TJ J XJ J XI TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ XI XJ TJ TJ T
CN CN CM CM CN CN CN CN CN CN CN CN CM CN CN CN CM CM CN CN CM CN CN CN CN CN CM CN CM CN CN CM CM CM CM CM CM cvi cvj cvj CM CM CM CM CM CM CM CM CM CM CM CVJ CM CM CVJ CM cvi cvi cvi
CD CD CD CD O) CD Ol CD CD CD CD CD CD 01 CD CD CD CD CD CD CD oi σi σi CD CD 01 CD CD CD CD 01 CD CD CD CD CD CD CD CD 01
O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O o o o O O O O O O O O O O O O O O O O o O O O
O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O o o o O O O O O O O O O O O O O O O O o ooo
CD CD oO
CDoO CDoO CD CD CD CD CD CD CD CD CD CD CD to CO CD CD CD CD CD CD CD CD CD CD CO CO CO CO CO CO CO CD CO CD CD CO CD CD CD CD CD CD CD CD CO CO CO CO CO CD CD CD CD CD CD
Ul LΩ LΩ LO LΩ Ul LΩ Ul LO l l Ul LO LO LO LO U LO LO Ul Ul l Ul Ul ) U Ul Ul Ul Ul Ul Ul U U Ul LO U Ul l Ul Ul LΩ U LΩ LO LO LO LO LO Ul U Ul U Ul LO U Ul CN CN CM CM CN CN CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM CN CM CN CN CN CN CM CN CM CM CN CM CM CM CM CM CM
(M CM CM CM CN CN CM CN CN CN CN CN CN CN CM CM CN CM CN CN CN CN CN CN CN CM OM CN CN CN CN CM CN CN CN CM CN W CD CO CO CΩ CO CO CD CO CO CO CO CD CO CD CD CO CD CO CO CO CO CO CO CO CD CD CD CD CO CΩ CD CO CΩ CO CD CD CD CO CO CO CO CO CO CO CO CO CO CO CΩ CD CO C
O LΩ O — LΩ LΩ LΩ LΩ CD CO st CO lΩ lΩ CD lD lD — sf CO CM LO sf LO CM CD LO CD S IO CM CO S CM Ul S CM CO CO CO CO OO OO OO OO sf O Ul S S CO Ul CO CD st S CO CD - sf Ul CO sf O Ul CO LO CM — S sf - S — — — CM — Ul — — st — — 00 — CM S — — — S CO — CO sf sf CM CD CO CN - — CO O OI S O S CD — O — 00 — sf OO O - CD O UI OO S CO — 00 — — — O — — — CsM CsN —s —s CsN —s CsM CsN CsN —s CsN OsCsCsN —s CsM CsN -s CsN CsM —s CsM CsN CsN —s CsO CsN —s —s —s CsM —s —s —s CsM CsN —s CsM —s —s CsM —s —s CsVI CsN CsM —s CsM —s —s CsM CsM —s Os —s —s —s —s CsM —s CsVI CsM CsM CsM CsM CsM CsM
— CN st sf LO lO CO S S OO CD CD O — CM CO CO sf CD Ul CD O CD S S S S OO OO OO O OO CO CD CO S S Ul Ul CD CD S O O O - CO CM CD CD CD O S UI UJ CO CD CD CO O — sf sf cO sf Ul CN 0000 00 C0 00 00 00 00 00 C00000 CD CD CD CD CD CD CD CD CD O CD CD 01 CD CD CD CD CD O CD O O — — — CM CM CM CN CM CO CO CO CO CO CO CO CO CO st cO st st sf st sf st Ul Ul Ul Ul CD CO CD S CD CO OO CO CO OO CO OO OO OO CO CO CO OO OO CO CO OO OO OO OO CD OO OO OO OO OO CO Or. I CO CD CO CD CD CD CD CD CD CD CD CD CD CD CD CD CD OO CD CD ai ro CD CD CD Ol Ol CD CD CD CD CD CD CD CD CD CJl CD CD CD CD CO CO CO CO CO CD CO CO CO CD CO CO CO CO CO CO CO CD CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CD CO CO CO CO CO CD CD CD CO CD CO CD CO CO CD CO to
CD »- s-
CM 00 I I CN H X CD CO I — rr LO oo I s <M T — CM I LO 00 I '- III I '- S CDl I ^IIII — S '- — '- '- I st l '- T T — — III cO CMl I I cO
S CN Ul CM CD O 00 sf sf X I — S 00 Ul , — X sf H CD X sf O S O I sf — CD sf l — CD st O — I S S OO LO LL CO l CO l I CO LΩ CO l st X X CO LO — CO LO CΩ CD CD —
S S CO sf CM S CD O f CD CD LO O S sf ro S CO sf CO O sf CD — OI CD LO CD — Ul sf CM O — CN sf CN O S CO Cvl CD CD Ul CD CO S CD CD O s on LO S OO sf CO O O - o
CO CN LΩ CD O sf S CD O CO CO f Ul 00 on s CD sf CD sf CD CD CD sf S CD O Ol — S CD S CD CM S — S σi LO CO O CD CD sf - S O sf CD LO o O CN 00 — O LO O CD CM
CO O CO S CO — CD sf OO CD sf sf CD CO CO sf CM CΩ C — sf CD CM C S sf — — sf cO CO LO Ul CD CN CO CD CO CD st - st cD LΩ CO S CM CD CO s lΩ LΩ CM CM S CO S S O
CN O O U CD LΩ IΩ O CD — — — Ul sf en st CD CD CD O sf — CM O CD — CN O CO CD O Oi σi O S CO - CO — CO CD sf CM — LO CM sf cO S r> ro CN CD CD O CO CO sf st sf
LΩ CVI sf S O sf LΩ sf CM CO CO CM sf CD o CO — Ul — OO OO — CM CO CO CO sf CD CM st st Ul CM CD CD st CD CM CD CO CD sf st st CD CO — — CM - CO S LΩ CN LΩ CN CD O CN
CD COCD CM CM CM CD COCM CDCD CD CD O — Ol CDsf CD t CO st CO — CD CDCM CDCM CO - LΩ CN CO S OLO CDS OI CN OlLΩ sf CD CN CD CO COCO CN CO C
O O O O O O O O υ υ o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o O O O vo φ φ φ φ φ φ φ Φ Φ Φ Φ Φ Φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ CD Φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ Φ o TJ TJ TJ XJ TJ XJ XI XJ XJ TJ TJ TJ TJ TJ TJ TJ TJ TJ XJ XI XJ XJ TJ TJ TJ TJ TJ TJ TJ J XJ TJ XJ TJ TJ TJ J TJ TJ TJ TJ TJ TJ TJ TJ XJ TJ TJ XJ XJ XJ J X) so CM CVJ CVJ CM Cvi CVJ CN CN CM CM CM CM CM CN CN CN CM CN CN CN CN CM CN CM CM CN CN CN CM CM CM CM CN CN CM N CM CN CM CN CM CN CN W 00 Ol oo CD CD CD CD CD Ol CD CD CD CD CD CD CD 01 Ol CD CD CD CD CD CD CD Ol 01 CD CD 01 f 1 o oo o o O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O OO O o o o o o o o O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O OO O
© to CD CD CO CD CD CD CO CD CD CO CD CD CD CD CD CD CD CD CD CD CD CD CD CD CO CD CD CD CO CD CD CD CD CD CD CD CD CD CO CO CD CD CO CD CD o LO Ul l Ul LΩ Ul LO LO LO LO LΩ LΩ LO U LO LΩ LO Ul Ul LO Ul U Ul LO Ul U Ul U Ul U) Ul Ul LΩ Ul U U l Ul Ul LΩ Ul Ul LO W LO LO Ul LΩ CM CM CM CN CM CN CN CN CN CN CN CN CN CN CM CM CN CM CM CM CN CN CN CM CM CM CN CN CM CM
CM CM CM CM CM CM (M CM CN CM CN CM CN CM CM CN CM CM CM CM CN CN CM CN CM CM CN CM CM CN CN CM ( CM CN CN CM CM CN C CO CO CD CO CD CD CD CD CD CD CO CD CD CD CD CD CD CD CO CD CD CO CO CO CD CD CD CD CD CD CD CD CO CD CD CD CO CD CD CD CO CD CD CD CO CD CD CO CD CD CD CD CD CD
U1 U1 U1 U1 U1 U1 CM — st CO CO CO CM st sf CD CM CD CO LΩ O O — — CD CO CN LΩ CJI UI O CO CO O S LO CO UI CO UI OO CO CO CD LO - sf CO CO
— — — CM — — CO — S CO CN O OO st S CM S lΩ lΩ LΩ — CO Ol OO st CD OO OO CO OO CD CN S O OO CO OO OO sf S CO O sf oO LO S — O1 C000 O1 CO CO CO — S O CO S OO OO OO OO OO CN sf C CN CN CN CM CM CM — — — CO CO — OO sf CO O CO OO Ol CD CM — — LΩ CO CO IΩ LΩ LΩ UI UI CO CO CD UI CN CO - sf - — CM CM CM CM CM CM lΩ Ul Ul Ul Ul lΩ lΩ st CO st cO sf UO Ul Ul CO Ul st sf S S S S S S CO CO CO CO CO st CO CO st sf cO CO CN CO Ul CO CO — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — vo
IΛ CO 0O 00 CO O1 — CD S S Ol CO sf - CO sf CD LO CD S S O O sf sf - CD S CD CO sf — — CO LO UI CD — S LO O O — — JN CO sf st Ul Ul CD CM CM CM CM — O O O O O — — CN CM CM CM CM CM CM CM S sf CM CD st CD CM CM CM CM CM - CM O — — LO IΩ UI IΩ LΩ IΩ CΩ LΩ CD S S S S
© — — — — — — O O CM O S S sf st sf st cO - CD sf O — Ol — — — — — — — — — — — — — — CO CO st sf st st O O O O sf sf sf sf st sf — — — — — — — — — — — — — © S S S S S S CO CD CO CD LO LΩ — — — — CD CD CD S — — Ol — — — — — — — — — — — — — — O CD CD OI CD CD — — — — — — — — — — — — — — — — — — — — — — — Ul
H U x_-xxxxxxir:icDiιιxιιι OO S O CNX" l oil *i- co — CO I I LΩ — ^ ^I coiII IIIIIcDsf l ooI st s I sf l l o co l u l oi l r:
CVl X O CD O CO CO CD CO l CO Ul CM S S CD CD CO CD LO CO - CD — CD CM CD CD OO COOOsf - OOOlX sX — COCO LOS OO CO CNCD sf cO O CD sf — — sf CD — (D CN CM S — LO 01 I LO Ol S sf S OO CO — — CD S CO CN Ul st Ul CO CO st st CM — O S LO OO CD CD S CM CM CD CD CM S CD CD CO CD S U CD S S CM — OOCO — Ul St OO LO OO O sf CM CD CD st CD CM Ul CO LO — OO CD S CD Ol CD CO CD CO LΩ st sf CM LΩ LO — CO CO CD sf O Ol S O CO lΩ CO Ul CO CO σiOl Ol OlCO— CO CD CO O S S COS sf LO SCO CO CM OO Ul CD st CN — S OO OO CO — S sf CM sf CD C sf cO CM Ul CO Ul sf Ul O - CO CM OO OO CD CO UI O CO O O — CO sf OO CD CO O — CD sf O! - sf CD sf cO CO O LO CM - — CM S CM CD CD CM O UI O S S CO CO OO OO CD S — — sf CO sf LO sf O - OI S S O CO S — LO CD sf — S CM O 00 S C0 — s — O O CO CD U — CO O O IΩ O UI CM CM — CO CM CO CO O CO — CVI CD — Ul — Ul Ul — CO CM CO CM O CD CO S CD CD CD — 00 C LO CD st O CD CN CM CO CM CD — — S CO LΩ f S - CO OO CM CM — CO CM Ol CM OO sf Ul CO OO - CO CO LΩ LO st CO CO S CO CO O st CO CO CN st CM CN CO LO LΩ S CM CD O — sf lO CO CN Ul — CO C CM Ul CN CO CVI st CM CO CO CD CO COCO CM CO CO CN CO CM OI CD CD CO COIΩ — Ol— CO CO COCO LO CD CDCD CO CO IΩ — — — CO Ul CO sf CM CD COCD CD— CO COCM CDsf CM CO COCD CD — COCD CD C o o o o o o o o o υ o υ o o o o υ o o υ o o o o o o o o o o o o o o o o o υ o o o o o o υ υ o o o o o o o o o o o o o ϋ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ TJ TJ TJ TJ TJ TJ TJ TJ TJ XJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ XJ XJ XJ XJ XJ XJ XJ XI T3 TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ T3 TJ TJ TJ TJ TJ
co co co co co co co co co co co co co co c co oo o co co co co co co co co co co co co co cri co co co co co co co co co co co co co n
CM CM CM CM CM CN CN CM CM CM CM CM CM CN CN CN CN CM CM CN CN CN CM CM CM CM CM CM CM CN CN CM CN CN CN CN CM CN CN CM CN CM CM CM CM CM
CM CM CM CM CM CN CO CO C CO CO O CO CO CO CO CO CO CO CO CO CO CO CO OO CO CO PI CO CJ CO CO CO CO CO C CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CD CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CD CO CO CO CO CD CD CΩ CD CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO Φ
S LΩ st S O LΩ lΩ Ul Ul lΩ CD Ul st Ul Ul Ul st S st Ul Ul Ul Ul Ul Ul CN Ul CO Ul Ul Ul CO Ul Ul Ul - LO sf S lΩ LΩ CD CD st CD CO O sf O — S CN IΩ LΩ CD CD CN UI O LΩ LΩ LΩ CD IΩ U UI — — — OO CN — — — — — — — — — — — — — — — — — — — — O — 00 — — — — — — — S — — CD — — — O — — CO CN — CM CM CO LΩ — — — — — — CM — — — — — — — C
CsM CsN CsN —s CsM CsVI CsM CsM CsM CsM CsM CsN CsM CsM CsM CsM CsM CsM CsM CsM CsM CsM CsN CsM CsM CsM CsN —s CsN CsN CsN CsO CsN CsN CsN —s CsM CsN —s CsM CsM CsM CsM CsM CsM —s CsM CsM CsM CsM —s CsM CsM CsN CsM CsM CsM CsM CsM CsM CsM CsM CsM CsM CsM CsM C
CO CO — — — CN OO CO CO LO CD CM st lΩ S S CO — CM CM sf CN CO CD CD CD OO OO OO CD CO CO st CD — CN CN CO S S OO CD CΩ LΩ CΩ S O S LΩ CD CO S CO CD — — S CD S S S S CO S CM OO C CO <r) CO CO CO CO (TJ C C C C st st st st st Ul Ul Ul Ul Ul Ul Ul Ul Ul Ul Ul Ul Ul Ul CO CD CD CD S S S S S S S S CO OO OO CO CD CO O O O O O O - CM CM CM CN CN CM CM CM CM CO CN C O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O — — — — — — — — — — — — — — — — — — ssssssssssssssssssssssssssssssssssssssssssssssssssssssssssssssssss
— - I
o o o
Cvi Cvi cvi cvi Cvi σi σi σi oi σi oi o o o o o o o o o o o o
C\I C\1 0ϋ C\J C\I C\J CΛJ (Λ] C\i C\] C\] C\I C C\J -\l -^ CD CD CD CD CD CD CD CD CD CD CO CD CD CD CD CO CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CO CD CD CD CD CD CD
OO CO CO CO — Ul CO CO S CD sf st CD CO CO CO S CD O OO st — CO S S t _
Φ N - r- cδ i N sf m » ffl D c» ω Φ ω » ffl s o sr ffl ffl o Φ » ro vθ θ θ _ st σi CO CO CO CD O O LO LΩ Ul st Ul Ul Ul Ul Ul Ul Ul CO st LΩ LO LO Ul CO Ul Ul CO st Ul Ul S CD CD sf σ OO S OO — O CM CM LO OO CD U1 CN - CM S CD CO CM CD — — CD CO CM LO CD CO OO S C CO CM CO CO CN CN CO CO CO — — — — — — — — — — — — — — — — — — — — — — — st sf sf Ul sf st sf CD CO st LO Ul Ul st lO sf LO S CD S OO CD CD OO - — LΩ LO CD CD CD CD CD CD C vo
IΛ JN OO CD CO st CO CO sf CD — Ul st st CD Ul st sf cO CD - O S S O
O O — — — — — CN CO CO CO CO CO CO CO CO CO CO st st st st Ul CD — S S CD CD CD CD O CN O CN CO S OO CD CD — CN CN CO CD O OO CD CO O CD S O sf LΩ CD CD C
© CN CO CO CO st -l LΩ CO LΩ CM CM CM CM CM CN CN — — — — — — — — — — — — — — — — sf Ul LΩ lΩ LΩ Ul Ul Ul CD CD S S S S S S S OO CD CD CD CD O S OO CO OO CO OO CD CD O CM CM C © UI LO UI UI UI UI UI UI UO — — — — — — — — — — — — — — — — — — — — — — — CM CM CM CM CM CM CM CN CN CN CM CM CM CM CM CM CM CM CD CO CO CD S S S CO CO CO CO CO CO st st st s Ul
H — — — CN — — CO CM O '"" s- — CM U I I I I I I I I I sf CO Z O CO O CD O CD CO l I H S CO — I S — 00 X S Ul I I Z O — II I I X X I I X I CM I X I II CD X X Z
S CD — sf CM CD CO LΩ lΩ CN CJl I — Ul Ol Ul sf Ul CO O CD CD sf sf — Ul I T I C LO — O CO CO CD CO CO CMIOICO I O CN X s CD f CM CD Ul 01 O CD CO 01 O 1 CD 1 sf r I: c sof I CD CD — CD U CD CO — I C
PH CO LO CM sf CM O CN CD Ol lΩ OO OO OO st CO CD CD CD — CD CD 00 CN — CN st CO — Ul CO CO — — — O LO CM CO CD sf CD s 00 CD sf Ol S sf 01 CO CM — CO CO CN CD CD O O CO CO CO CO CO CD S
CM st CM CM Ul CN O CM CD Ul CO Ul sf sf cO O CM CD CO O S CD CO — S CD OO O S CN OIOLΠ- st sf CD CO CO Ul CO — CO O LO O CO CD 0000 — CO CN CO sf CD CM CO CO
CO 00 sf CM S lΩ CN S CO — O O CD CO CO S — sf — U S CD CM CO CD LΩ CO 01 S LO CN CO CO Ul S S LΩ CO sf O Ul CD CO Ul CM 0000 — CO LO S CO CD CD CO — CO 00 CM O — S O sf
01 Ul CD Ul CO — σi LΩ CM S CN OO CO Ol CD CO — st Ol OO CO CD LO O CM Ol CM CO O — CN CO — LΩ CO CO CD CO st CM st st sf O CO CO — — 00 — LO sf sf CD O — O LO O tD O CD CO — CO L
LO sf CO O — — OO sf - CM CO OO LΩ CO LO LO st cO CO CO O sf sf CO sf O CO sf sf Ul CM CO 00 f Ul LΩ CO — sf CM CD CO CO Ul U S sf sf OO sf CM CD S CM CO CM CD sf CD CO 00 — CO 01 CD C
CO CN CO CO CN CO — N CO CO COUl CO CD CD CO CO CO CO — sf — CD CD CDLO CO O) COsf CD CDCM CM CO CO CO S CM COS Ul CO CM CO — CO — — — — CD — CM — 00 COCM CN U) CM CO COCM sf 00
O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ TJ TJ TJ TJ XJ XJ XJ TJ TJ TJ TJ TJ TJ TJ TJ TJ -α TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ
CM CM CM Cvi cvi cvi cvi OM CN CN Cvi cvi cN CM CN CN Cvi cM CN CM
Cvj Cvi cvi cvi cN Cvi cvi <M CN Cvi cN CN Cvi cN cvi cN Cvi <M CN CN CN W CJl CD CD O) CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD a) CD CD 01 Ol CD CD CD CD a) CD CD CJl CD CD CD CD CD CD CD CD CD CD CD 01 Ol CD CD CJi eD CD O) oo oo co oo co ∞ co co co co co co co co co co co ∞ αi oo co co ro oo co rø ∞ co co co co co co cri c co c co co co co co co co co co co co co co co co co co co co co co co co co co co co co co co co co n
CN CN CN CN CN CM CM CM CN CM CM CN CM CM CM CM CN CN CN CM CM CM CM CM CN CM OM CN CM CM CM CN CN CN CN CN W co co co co co co o co co co co co co co co co co co co co co co co co co co co co co co co co co co co co co co co co c co co co co co co co
CD CD CD CO CO CD CO CO CO CD CO CO CO CO CD CO CD CD CD CO CD CO CD CD CD CO CD CD CD CO CD CD CD CD CD CD CD CD CD CO CO CO CO CD CD CD CD CO CD CD CD CD ω t
X) ca H
CO S CO CM CD LO — LO CD CO CO CD CN S LΩ sf CO CM LO S st CO st LO CO CO CO O CO O CN CO O CO O CO CO — CO LO CO CD Ul Ol O Ul sf CO - CO CD CO CD CD sf CO S CM CO sf cO CO CO OO
» M ffl ro ιn Φ θ f ffl o m s o m _ ffl ω w Φ co _ ffl Φ f ω ffl ω ω ω _ ω » ffl
LΩ LΩ Ul LΩ Ul CO Ul sf Ul CM IO CO CO Ul lΛ lO sf sf st LO Ul in Ul Ul Ul Ul lΩ LΩ LO Ul UJ UJ Ul Ul Ul Ul Ul lO IO Ul LΩ Ul Ul Ul LO Ul Ul Ul Ul Ul Ul Ul lΩ LΩ lO W
— — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — CM s
CN sf lΩ S O — CN C C CO σ — CN — Sf CO — S OO S CD CO — CO sf O S — LO — O CM S OO LO CO CD CN st — CO CM CO CO CM CM st OO CM LO CD — OO OO OO — — S CM CM CO CO CO O S S S S OO CO OO OO OO CO OO CD CD CD CD CD O O O O O CN CO CO CO sf C sf st Ul CD Ul LO LO CD CD CD S S CD S OO CD OO CD CD CD CD CD CD CD O CD CD CD O O O — — — — — CM
— — — — — — — — — — — — — — — — CN CN CM CN CN CO CO CO CO CO C CO CO CO CM CO CO CO CO CO CO CO CO CN CO CO CM CO CO CN CO CO CN CN CM st CN CN W
— — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — CO CO s
— CN — — — — — CO — CD st 00 CN CD CO X X — LO X sf CD X CD — f II I 00 CO CM H I I lΩ CN st I I ;- _- H S CD CN LΩ 00 O 01 O I CO — CO II CN CO ΪZ st l cO CO — X I oO CO OO CD S l I
CM CN CO X O CO S CD CO S LO St 00 o s sf O CO sf st S CO sf CN Ol I I CD CN CO CO OO st CD — OO CO S — O S O X S O l O OO Ul Ol OO CD CD S Ul S S — CD CM C
CD on CO CO s o CD s O CO CO s s CD Ul CO CD s CD O 00 CO CD CM — 00 — CO O sf CO CO CO CM CM LΩ CD CN CM — CO S CO CN S CO CO CO — CM O CO CD CD S st LO S CD OO CN Ul CM CD C cn CO to s CM CN CM CN sf CM LO , — on on CD sf LO LΩ CM LΩ CO C αOj CDO CvDD Ov_) OC0O UlΩI SlS Ov-? CvO) OSJ -τ- —τ- σCJi) COJl) CCO) -τ- LLOϊ; ssfr C,O) Ov_) CD CD S CD CM CO CD CO OO — sf sf O sf OO - CO CO CO S S
CO o sf CO s O O LO CD CD 00 CO CO sf O , — 01 st CVI LO St O O CM CO CVJ S CO CO S st — LO CD CD LO CN CD — — CM CO st c CbO CCMN SS Oσii uUil sstt oooO cCoO -— O O ULO ssff o sf O s , — LO CN O 00 00 CM sf CO CO o S ro CD Ul on o LO S O CM CN CM LΩ LΩ LΩ OO O IΩ IΩ CD CN OO LΩ — — O sf CO CM CO CD sf — S OO CM CM O CD sr oO CD — — OO CM CO CD C sf LO st CO f CO CD CO O sf CN sf CO LO CM s O 00 CM St CO st CM CO CN sf - CO CO O OO CO LO UI CO CN CM CD CN — S CO CO CN CM sf st — C — CD — CM — O OO — — CN CO CO LΩ sf
CD CD CD CD O CN CD CO CD CVI CO COCD CO CO COCM — sf CO CO CO COCN sf LO CD CO COCM CO CO CΩ — C0 C0 C0 CD C0 C0 C0 C0 C0 C0CO C CO COUl CO CDCD COCM CO CO COLO CN CO CO CO CO COCO — o o o o o υ o o o o o o o o o o o o φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ o φoφoφoφ oφoφoφoφoφoφoφ oφoφoφoφoφ φooφoφ oφoφ φooφoφoφoφ o φ oφoφoφoφoφ oφ oφ oφoφ oφoφoφ oφoφoφ oφ oφoφoφ oφ oφ vo XJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ XI XJ XJ XJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ XJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ XI TJ T o CN CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM C CM CM CM CM CM CM CM CM CM CM CM CM CM tM CN CN CM CM CM CM CM CM CM CM CM CN CM CM CN CM CM CM CM CM CM CM CM CM CM CM CM CM CN CM CM CM CM CM CM C so
CN CM CM CM CN CN CM CM tM CM CN CN tM CN CN CM CM CM CN CM CM CM CM CN CM CM eN CM CN CM CN CM CN CM W CM CM CM CM CN CM CM CM Cvl CN CN CM CM CN CM CM CM CN CM CM CM CM CM CM C
CJ) CJ) tD <D CD CD CD O O CD CD CD CD CD CD O CD O CD vD vD vD vD CD vD vD vD vD CD CD CD CD CD CD CD CD CD CD CD CD O! CD σ) σ) σ) σ) θi σι σ) σ) σ) σ) θi cD CD CD CD CD σi σi σi cD cD σi σ cD C
∞ CO OO CO OO CO CO OO OO CO CO CO tO CO OO OO CO OO CO CO OO OO OO OO OO CO OO CO tO OO OO CO CO OO OO OO OO CO CO OO CO OO CO OO OO OO OO OO CO CO OO OO OO OO OO OO CO OO OO OO CO OO OO CO OO OO O
© o co co co co co co co tri co co co co co co co co co co co co co o c co co co co co co co co co co π co co co co co co co co co co co co co co n
CM CN CN CN CN CM CN CN CN CN CN CM CM CN CN CN CN CM CM CM CM CM CM CM CN CM tM CN CN CM tM CN CN CN CN CN C co co rt co co co crj co co trj co co co co co co co co co co co co co crj co co co co co tri co co co co c^
CD CD CD CD CD CD CO CO CD CO CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CO CO CD CO CD CD CD CD CD CD CD CD CO CD CD CD CO CD CD CD CD CO CO CO CO CO CO CO CD CD CD C^
CO LO CO CO CO — sf S sf cO OO CM CO CO S CO CO CO CO CO CO CO lO sf cO CO CO CD OO O CO — S sf cO CD — CO CD O CD S LΩ CD CO sf cO CO sf UO O O CD CO O sf S CD Ul Ul s
OO CO OO OO CO CO CO OO OO CO OO CO OO OO S — — Ul LO CM LO CO OO OO CO OO CO CO CO OO OO CO st CO CD CD CO - OO CD st CO LO S OO st cO CD sf CD S sf st σJ CM sf CD LΩ CO OO sf - sf S CD S C LO Ul Ul Ul Ul CD LO Ul Ul Ul Ul Ul Ul Ul Ol CO O CO LO S st Ul Ul Ul Ul Ul Ul LO LO LO Ul Ul Ul Ul sf Ul Ul st LO sf - — — — O — — — — — — — CO O — — — — CO CM CO CO CO LΩ CO LΩ C
— — — — — — — — — — — — — — sf LO LΩ LO LO Ul - — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — vo
IΛ JN CD CD — lO st S Ol CO σi S CO LΩ O — sf sf st LO CO S O Ol O LO CD S S S O CN Ul S OO S S CO CO CD IΩ CO CD O C
CD CD S S S S S OO CO CD O — CM CM — CO LO CD sf CO — — — — — — CM — CM CN CN CN CN CN CO CO CO CO CO CO CO OO OO CN UI S LO — CM CN CM CO CO CO — CD CO CD CD CD CD CD S S OO CD C
© st st sf sf st sf st sf st st LΩ Ul LO Ul OO CO CO OO CD - CN CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM CO CD CO S OO OO O — — — — — — — CM CM CM CM O O O O O O O O © — — — — — — — — — — — — — — CM CM CM CN CM CO — — — — — — — — — — — — — — — — — — — — OO CO OO OO OO OO CD OI CD CD CD CD CD CD CD CD CD CD — — — — — — — — Ul
H CO U CO I CO CO C S CO sf sf O O CO O H .N s itoiri 11 = = i CO T
S CD CN CD CO I CO S I I — O S sf — O — CO Ul — CN CD S Ul CO — S CD I I I I I I
— CO CD Ol — CM CO CO CM CD OI CO O CN S CD CN LO W J I.c Ioco s |vo uoy — r- _ l o gOj uOsf — — I __ I _ι_ OOiOOiCDiO COLC-OXCNxOlxCDxUlxOOxstrl:! — I I 01 I s
CD sf CD S 00 CO — Ul S S sf Ul S 00 CO S LO S — — sf C0 O1 O1 — OM OO S CO OO OM OO CM OO OO S S CO CO S UI CN CM UI O S CM — lΩ sf sf sf st S - 01 — 00 00 — sf CM CO Ul CO CD
00 — sf 00 CN S O CD sf O CO — CD Ul CO U CD CO sf CO CD CD CO O sf CD CO LΩ sf S sf CD CN sf O CD CD S st O CM S - CO CO — O st CD CM CO sf st sf O CD LΩ CO CN — CM — O CO 00 01 C
CO — S sf sf U — O O CO CD Ul CO Ul — O Ul CD CM CO O S CO CM CD S sf Ol LΩ - sf CD U) CO CD sf CD OO LO CD CN - — CO O S — 00 — S O LO LΩ LΩ — OO CO OI O O — st O CO CO CD
O sf 00 — O st — O Ol CM CD CN CO O CO CO CO CO O Ul Ul O sf CD CM CD CN CO UJ CD CD CN S O S CD CO CO S CO σi S LΩ CD CO CD OO — lΩ CO sf Ul Ul S Ul CD CD — CD CD — CM S Ul S OO C
CO CM sf CO CO S CO CO Ul O — CM st CO CO S U CO CD O O st CO — sf Ul sf CO Ul CO CO sf cO st st sf cO CO CN CD — LO CD S U CD — O CD Ol CD CD CD st S Ul Ul CO OO O st st — CD 00 CM s
CD CJ) O) CD Ol CO CD COCO CD CO CO CO CO— — CO CO CO CM CN CD CD CD CD CD CD CD COI CD CD CD- COCO CD— CD CO CDCO — CO CO CN — CM CM CM CM — — CM CM — sf sf CO Ul 00 CM CD CDs oφ oφ φo oφ oφ oφ oφ oφ oφ oφ oφ oφ oφ oφ oφ oφ oφ oφ oφ oφ oφ oφ φυ oφ oφ oφ φo oφ oφ oφ oφ oφ oφ oφ oφ oφ oφ oφ oφ φo υφ oφ oφ oφ oφ oφ φo oφ oφ oφ oφ oφ oφ oφ oφ oφ φυ oφ oφ oφ oφ oφ oφ oφ oφ oφ
O O O X O O O O O O O O T| X; O O O O X O O O O X X O O X O O O O O O O O O O O O O O O O O O O O O O O O T O TI O O O TI t| O t) O O X O
CM CvJ CM CN Cvi tλi cvi eM CM CN CM CN tM CM CM CM CM CM CN CN' lM CN CN tM CN
CM CN CM CN CN CN CM CM CM CN CN CN CN CN CM CN CN CM CN CM CM CN CN CN CM CN CM CM CM CN CM CN CN CN CN CM CN C CD CD CD CD Φ CD CD CD CD CD CD CD CD 01 CD σ) CD CD 01 Φ Φ Φ Φ CD Φ CD CD Φ Φ Φ Φ CD Φ Φ Φ Φ
∞ oo αo ∞ oo ∞ αo co oo αi αa co co co αj oo oo co oo oo oo αD oo oo oo oo co ∞ co co co co co co co co c co co co c co co co co co co δ co co δ co δ δ co t i co n tN tM CM CM CM CM CM CM CN CM CM CM CM tM CN CM CM CM CM CM CM CN CM CM CN CM CM CM CM CM CM CM CM CM CM CM CN CM CM CM C^ co tri trj co trj co co t co co co co co co co co co c co co co co co c c co co co co co co co co co co co co tri co
CO CO CD CO CO CO CO CO CO CO CO CD CO CO CO CQ CO CO CO CO CD CO CO CO CO CO CO CO CO CD CO CO CO CO CO CO CO CD CO CO CO CO CO CD CO CO CO CO CO CO CO CO CO Φ
X) c3
H
CO CO CO OO CO CO CN — CO CD S OO sf — CO CO CO O CO CO S st cO CO s CO LO OO Ul S OO sf CO Ul st — CM OO O CD S O CD CD LΩ OI CD O OO O CD O — CD CD O O CD OO CD Ol CD CO sf — S O OO CO OO OO OO CO CD CD CO S CO Ul OO CD CO OO OO OO OO CO CO CD OO st O S CM O CO S CO S S CD — Ol sf Ul lO sf — st CM — CD — CD CD CM CO CD O CD OI CD O CD OO UI CO O — CD — st cO CO lΩ Ul Ul Ul Ul Ul Ul Ul Ul CO LΩ LΩ Ul LO l Ul Ul LΩ LΩ LΩ LΩ Ul Ul Ul U CΩ S S CD S CD S S S S S CO S CD CO CD CD CD CD LO LΩ CD CD CD LO LΩ CD st LΩ LΩ CO CD st lO CD CN —τ~ C CM " —
CM CO st CD CD 00 s CO o CO 00 CD CO CD CO CO S st O CD LΩ Ul
CM CD sf st Ol OO CO CO CD CM CD S S LΩ CD S O O st CO CO CO CD lO CO OO CD CD CO CD CO sf - — sf 00 00 00 00 CO 00 OO CO 00 CO 00 00 OO 00 CO o o CO sf sf sf LO Ul C CO LΩ CD CD CD S CD CD CD O LO CO OO CO CO CO S CD — CN CM CM CM CO st sf sf sf Ul Ul CD CD S S S CN CM CM CM CM CM CN CM CN CN CM CM CM CM CN sf sf sf sf sf sf sf st sf s sf st sf st sf st st sf st Ul Ul Ul Ul vO CD CO CD CD CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO — — LO sf CD
— — CD '-
I I I I I I I I I I I I l_ I I - O S CO CN S LO O CO CO LΩ S — — S CO LO I I r o tD s = '- c
CD CO sf LO s CN 00 oo LO CD LO LΩ CM 00 U I I I I I I I cO Ol I O l I I I I I CM ∑IZ Isf I CO Isf Xr: L 00 CD ICD ICD sf O S CO OO sf CD O O CO — CVI CD — LO ~' — -" -'> <" " — O — sf Ul ro ,— o CO sf CM CM S CO LO CO LO — — — CO S CO CO sf CD LΩ S 00 CO S CN S CO — CM S CD OO LO
CO CO CD Ol s s U) O s l CVI CM CN — CO CD CO sf LO CD 0000 CO LO CD sf Ol LO CD CO CM Ul S sf CD O CM sf - CO S LO LO st O — S S
CD Ul — CM CO sf U CD st CN LO LO O sf 00 sf — CO S — CO 00 — LΩ CO 01 LO — CD OO CJI CN S CN O — CO CO O sf O OO OI LO S CO CO C
S CM S CO CM on CD Cl st s 00 S S CD S — sf CM LO Ul sf — O — sf O S 00 S CO LO OO S S CD — O S O S O O CO CD OI CN S O S S O CD — CO CO CN S — sf SS CCVJ CCOO -— CCNN CCOO CCOO
S 01 CN CM CN st CO LO CO CD CO CO CO — Ul — CO O S LΩ CO CO S CM Ul sf S CM S CM O — — — CO CM CO Ul st oO CO CD CM sf st cO CM sf - — st CM sf CM sf UO CM CO CM OO st st st — CO OO C
CM CM — — CO — CO CO O — CM — — CO — OICM sf CO CM — S CD CDCM CM CM LO CN CN CM CO CD CJICM C CO CO CM — — CD CD CD CD CO CO CO CD CO CD CD CO CO CD CD CD- CN CO CD CD CD CDCD CD o Φ oΦ oΦ oΦ oΦ oΦ oΦ oΦ oΦ Φυ Φo oΦ oΦ oΦ oΦ oΦ oΦ oΦ Φo Φo Φυ oΦ oΦ oΦ oΦ oΦ oΦ oΦ oΦ oΦ oΦ υΦ oΦ υΦ oΦ oΦ oΦ υΦ υΦ oΦ oΦ oΦ oΦ oΦ oΦ oΦ oΦ oΦ oΦ oΦ oΦ Φo Φo oΦ oΦ oΦ oΦ oΦ oΦ oΦ oΦ oΦ oΦ oΦ oΦ oΦ vo TJ X! TJ TJ T3 TJ TJ TJ T3 TJ TJ TJ TJ TJ XI TJ TJ TJ X! X! X! X! T3 TJ TJ TJ TJ T3 TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ o CM CM CM CN CM Cvi cvJ CM CM eM Cvi cM so
CM CM CM tM CN CJ CN Cvi cN CM tM CM' CN Cvi cM CM CM CM CM Cvi tλi w
Φ CD Φ Φ CD CD Ol CD Φ CD OJ OJ CD Φ CD CD Φ CD Φ Φ CD Cfl CD CD Φ Φ Φ Φ CD Φ Φ Φ vD Q co ∞ rø oo oo co oo oo oo oo oo oo αi co ∞ rø ro oo oo co co co oo co ∞
© o CO CO CO CO CO CO CO CO CTJ CO CO CO CO CO CO CO CO CO W CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO C CO W CN CM CM CM CM CM CM CM CM CM CM CM CN CM CM CN CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM CN CM CN CN CM CM CM CM CM CM C co co co m co w co co c co co co co co co co co co co co co co co co co co co co co co co co co co co co co co co co co co co
CO CD CD CO CD CD CD CD CD CD CD CD CD CD CD CO CD CD CD CD CD tD CD CD CD CD CD CO CO CD CO CO CD CO CD CD CD CD CD CD CD CO CD CD CD CD CD CD CO CD CO CD CD CD CD CD vD CD CO CO
Ul CO Ol CO CN CO CD Ul sf CD CM CM Ul CD OO OO - sf O sf - — CO CD Ul CD Ul CO Ul Ul LΩ Ul LΩ Ul st v^ sf Ul Ul st M st Ul Ul Ul lΩ Ul LO LΩ Ul st Ul CO Ul LΩ LΩ CO Ul Ul Ul CD CO CO CO CO - C CO sf CO sf st CN st - sf CO LΩ sf sf lΩ CO st sf CO st sf Ul st st st st Ul - 1- — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — Cvl CM CM CM CM vo
IΛ CD CO CD — — CO CD CD CD CD O — sf Ul st Ul S sf sf st CD CO S CD CN CN CO CO - UI S CO O CD CD CD CD CO UI CN JN CN LO CD CO CD CO CN CN S CD — Ul O CO O CO CO sf CD O OO CD — OO CO OO CO CO CO st st st st st st st Ul lΩ LO LΩ Ul lΩ Ul CD CD CD CD CO CD CO S S S S OO S S S OO S CM S S OO OO CD C
© sf sf LΩ CD CD S CO CO OO CO CD CD O CD O O - (M CM CO CO CO sf sf sf st CN CM CN CM CN CN CM CN CM CN CN CN CN CN CN CM CM eN CM CN CN CN © — — — — — — — — — — — — CM — CM CN CM CN CN CN CM CN CN CN CM CN — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — Ul
H U I I CM X X X X co I I CD TZ - III CNl I I I rrloooocoi S 01 I I — I LO O CD I CO I O — LO I I col X OO Ul I — O 00 I CD I sf CN CD 00 I
CD O S — CO 0000 CD CO — — st I CN CO CO — O sf — o — S CD LO CM S CO CO 00 CO CD sf — CO 00 CO O CD S CD S S S t CO sf O — CO CO sf - CM CD CO CO CO — CD S 00 sf CD CD C
PH 0 Ul 00 S CD Ul LO CO — s CD CO CD — CO CO CO st S OO st CΩ LO S CO — S CN — S CN 00 CO CD CD CD CM sf CM 00 sf Ul sf — 00 Ol Ul sf O CD O CO Ol S S OO sf Ol sf oO st st O LO CD Ul C
01 S O CD Ul O CM sf CM CD CO CM Ul S CO S LΩ CO OO S S CO co co σi co ui σi co co sf CD CM — 00 sf CD 0001 CO sf CM sf Ul O Ol S O — S Sf O ~ CO s O1 00 C0 CM CO 00 CO O O — O S 0 S Ul sf O LΩ U LO CO CD CN O S 00 CD — lO O CO CD sf CO OO sf OO CD O S Ul - CO 010100 — CM S — — S — 00 CO CD CD O 01 S sf Ol st S CD 00 CD S CM 00 CD CD CO CO S sf — — sf C — 00 CM 00 U 00 U CO Sf CO CN st 00 CO Ol — sf S CO st CO S OO U CN CO — CD — CD — CO Ul s st CM CD 01 CM U S CM CD CM O Ul CM LO CO CO — 00 00 S — CD S O O S CD CO 00 CO M Ul C CD CO CM — CM S st CO 00 LO — 00 S — CM CD sf — — st lO CD sf lO CD - sf CN sf CM sf CO CD CD CO 01 — CO CO U CM sf CO 00 CO sf 00 00 st CO O st cO OO lO st CM CM O CM sf CO CO sf CM CM C CM CD COCO CO — CO COCM — CO CO 00 COCM CM sf COCO — CO LO CN 00 LΩ CO CO CO COCD CO CD— sf CO— CO CD COsf COCN CO CO CO COCM Ul CO— CM CD COCO CO CD CD CDCD COCO CO CO CO COCO s o o υ o o o o o o o o o o o o o o υ o o o o o o υ υ o υ o o o o o o o o o o υ o o υ o o o o o o o o o o o υ o o o o o o . —. —. τj χj χj χj τj τ τ τj TJ TJ TJ TJ TJ TJ TJ TJ Tj τ τ τ -q -q τj τ χ^ o o o o o o
CM CM CN CN Cvi cvi tM CM CM Cvi cM CM Cvi cM tλi cvi tM CM CM Cvi cN Cvi CM' w °. °. °. °. °. °-
CN CN Cvi cvi cvi cvi cM CM Cvi t CM' CM CVJ tλi eM CM CM CM' cvi tM CM' cvi CN cD tD σ) CJi cD cD cD σ) σ) CD CD cD CD CD CD CD CD CD σι oι oι oι σ) σ) σ) θi oι σ) σ) σ) σ) σ) CD CD CD CD CD CD σι σ) σ) σ) σ) CD CD CD CD CD CD σ^ o —o c —o c —o c —o o —o c —o c —o c —o c —o t —o c —o c —o o —o —co c —o o —o c —o c —o o —o c —o c —o c —o c —o —ro o —o c —o c —o c —o c —o —oo c —o c —o —oo o —o c —o o —o o —o o —o o —o o —o c —o c —o c —o c —o c —o — — — — — — — — — — — — — — — ssssss co c coco coco co co co co co co co o coco co coco co co co co co cococococo co coco co co c co trj cόcoco co co coco w
(M CM CM C CM CN CM CN CM CM CM CM CN CN CM CM CN CN tM CM C tM CM CM CN CN CM CN CN CN CN CN CM CN CM CN eN CM co c n n n co n n n n n o n c co n co o co n n n c co ci co ci c o to co n
CO CO CO CO CO CΩ CΩ CO CO CD CO CO CO CO CO CO CO CO CO CO CΩ CD CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CΩ CO CO CO CΩ CO CO CO CO CO CO C^
X) C3
LO S Ol CO O sf Ol sf — Ul S CO sf S OO CD O - CD sf cO CO — O CN U1 O S U1 S O CD 00 O
CO CD S CM CD OO CO CD CO CM CD st CD CM — CM st CD ω cO O st sf cO CM CM sf — LO Ol — S CO OO O OO sf CD Ol CD sf — OO CD S CD O O Ol CN CO CM CO lΩ OO sf OO CM CO O OO CO — — O O s CO CO UI CO IΩ LΩ CO CM CM CM — UI CM CN CM CM CN CN CN CM CM U CO CO CO CO CO CO — CM CO — CO IO UI OI IΩ OO CD - — LΩ S CM OO — O CN OI O OO OO LΩ - 00 — S CM O CM CD O — O CD S 0 — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — CO CM CO C CM CO CO CO CM CO CM CO CO CO CN LO CO CM CO — CN CO CN CO CO CO CN CO CO CO CO CO C
CO CO CO O S CM LO sf Ul sf Ul CD Ul Ul Ul Ul Ul Ul Ul Ul CD O CM CM S OO OO — LO S S O CD CM σi σi Ol O O O O CM CO CO CO CO CO CO CO CO CO CO CO CO CO st st sf sf st st Ul Ul Ul Ul CO Ul CO CM CM C
O O O — — — — O O O O O O O O O O O O O O O O O O O O O O O O O O st Ol S CD CO sf LO st cO CO CD CO S S S S S sf O S CO O CD CD CO S O CD CO — S sf sf s — — — — — — — — — — — — — — — — — — — — — — — — — — — T- — — — — — — Ul Ul Ul LΩ lΩ LO LO lΩ LO Ul Ul Ul Ul Ul Ul Ul lΩ CD Ul LΩ CD LΩ LΩ Ul Ul tD Ul Ul CD Ul — —
r:ico cDi :i I I I I I I I I I I I I I X I I I I I I Γ- I I IIII
I CD CDiCOsOf CD —ι —ι OισiiCOHCDiCDrl UlιSιCOιOιCNιOΪCMHCOi —ι —ι OϊOOϊOlrI:iOιOOi —r I:ϊsf ιCOιsf S OO sf u o oi CO — sf S Ul 00 00 I00 XO ICO s sr Ul sf oo Ul LO CO l Ul CD sf Ul CM - C OO CO CD CO CD — CD S st oO CD Ol CD Ul CD CN LO CN sf — CM CD CD Ol st CN S CM LΩ st st O Ul CO CM CO — Ul CD S O CO 00 O CD O sf CO — — sf s s CO CO CD , — CO CD CD sf CD sf st O L — CO CD CO σi CO S CN sf S CO LΩ O — O CM CM LΩ CM CO OO st CD st S OO CO O σi S OO sf CO st cO 00 — s o o s U CD Ul Ul CD O CM 00 O ss st LO ff s s S CO O O CO CD CD S CO UI CO LΩ UI C sf O sf - OO S CD OO OO CO S O CO CD O O O CO O CD CD CD - CM Ol O OO CO CD Ul st CO S CO UJ CO CN S sf 00 CM — U CD CO CO CD S CN sf o 0n0 CCDD s CO , — S — IO CD IO CD CM CO O CO CO CD CM — CO S CO CO LO Ul Ul CO CO Ul Ul CD st Ul OO OO O O OO st CM CD CD O st — S — S CO CD CD sf CO S 00 sf U CM st CD CO CD CN CO s l CO 01 CD — — CD CN sf Ul CO sf sf — — CO — LΩ OO OO — S Ul αO OO CO S CD CD Ul Ul S — CD CO CO O CO — sf st CD CN Ul Ol sf UJ sf CD 00 CO CN sf O CO 00 Ul O sf CM sf — sf o oo st CO s CD CO CO CO LΩ CN CD CO CO LΩ CO COCO CO CO— — CO — — — — CD — CM CM st CO CN LΩ CN CO CO CO LO CO CO — Ul sf CO - — CO CO CM — CO CM — CO — CM U) CM CO CO CD CO CN CO — CM — — CM — CO — CM CO — CO — C o o o o o o o o o o o o o o o o o o o o o o o o υ o o o o o o o o o o o o o υ o o o o o o o o o o o o o o o o o o o o o o o o o o o φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ vo TJ TJ XJ XJ XJ XI TJ TJ TJ T3 TJ TJ XJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TI TJ TJ TJ T3 T3 TJ TJ TJ TJ TJ T3 TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ o CM CM CN CN CM CM CM CM CN CN CM CM CM CM CM CN CN CN CN CN C CN CVJ C so
CM CM CN CM CM CM CN CM CM CM CM CM CM CM CM CN CM CN CN CM CM CN CM C CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD Ol Ol tD tD Ol Ol Ol C co co rø co oo oo co co co coαa oo co co oo αD cococo co oo oo co coco rø OO OO OO OO OO OO CO OO CO OO OO CO 0O CO CO CO 0O 0O 0O CO 0O 0O 0O C
© o co c co co co co co co co co co co co co co co co co co co co co co co co co co co co co co co co co c co co co co co co co co co co co co c^
CN CN CM CN (M CM CM CM CM CN CM CM CM CM CM CN CM CN CM CM CM CN CM CN CN CM CM CM CN tM CM CN <M CN CM CN CM CM W co co co co co co co co co co co co co co co co co co co co co co co co σi co co co co co o cό co co co tri co co co co co co co w
CO CD CD CD CD CO CO CD CO CO CO CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CO CD CD CD CD CD CD CD CD CD CD CD CO CD CO CD CD CD CD CD CD CD CD CD CO CO C^
CO O cO S C CM — C sf CM C CO CD
xsf O U CD st CN st o o o o o o o o o o o o o o υ υ o o o o o o o o o o o o o o o υ o o o o o o o o o o o υ o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o
UI U U LΩ LΩ LΩ IΩ U UI IΩ UI U UI UI UI UI UI LO UI IΩ IΩ IΩ LΩ LO LΩ LO LΩ UI UI UI UI UI UI UI U UI IO LΩ LΩ IΩ LΩ U LO UI UI UI UI UI IO UI LO UI IO LΩ UI Λ sf st st sf st sf sf st st sf sf sf st st st st st st st sf st st st st st st st sf st st st st st st sf st sf st st st sf st st st st st st sf sf st sf st st st sf sf sf st S S S S S S S S st st st st st st sf st st st st sf st st st st st st st st st st st st st sf st st sf st sf sf st st st st sf st st st st st st st st sf st sf st st sf st sf sf st sf st st S S S S S S S S
S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S OO OO OO OO OO CO OO OO O
CD CD CD σ) σi σi CD CD CD CD CD CD CD CD CD CD CD CD σi σ) σ) σ) σ) CD CD CD CD CD CD CD CD CD CD CD CD CD σi Oi σi 01 0i σ) σ) σ) CD CD CD CD CD σ) σ) σ) CD σ) CD CD CD CD st st s^
— — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — co co co co co co co co c st st st sf st st st sf sf st sf st st st st st st st st st st st st st st st st st st st sf st st st st st st st st st st st sf st st st st st st st st st st CD CD CD CD CO CO CD CD CD CD CD CD CD CD CD CD CD tD CO CD CD CD CO CO CD CD CD CD CD CD CD CD CD CO CO CD CD CD CD CO CD t
— 3
LΩ st CN CO O sf CD CD LΩ CD S CD — — LΩ OI CO CO O CO CD — S CM CD S S — CM — OO CD — CD CD — — CO st Ol — O — O OO LΩ sf cO CO — O S CO Ul CD S CO OO Ul O CD CN S sf — LO CD O CN st — LO Ul CO OO CN CD st CM lΩ st oO OO CM st st CM OO O - CD CD — CD CD CO — S CO CM CM — O CD — CM CM CM CM CM — OO C CD LΩ CD LΩ LΩ U OO CD OO O CM S CD O OO CO CN OO CO - O O LΩ CO LΩ LΩ LO LΩ CM CO — CO CO CO CO — CN CM CO — CO S — CM — — CN S S LO S CN CO CO CO CN CO CM CO — CO CO CO CO CO CD C — — — — — — sf sf sf LO — — — — — CM st st CO CM CO st LO sf Ul st st st CM CM CM CVl Cvl CN CN CN CN CM OO OO CO OO — — — — — — — CM CM CM CM CM CM CN CN CVI CM CM CM CM CM CM CM —
CM sf O Ul 00 CM 00 CO O Ul U CO OO CD O CM sf Ul - CD — 00 st st CM S CO OO CO CO CO sf S OO OO O CO lO CD CO Ul C st S sf LO CD O sf — st sf st CD CN O sf Ul OO CM CO CD O O — CO CD S OO OO CD CD CN st CD lO st CD CN CN st — — S CO — CM CD CD CD CN CO sf st OO OO OO O st — C CO CO st C CM CO O CD CD CD CD CD O CD st O O LO LΩ CD — CD OO OO OI CD CD CD CD CD OI CD LO CD S CO O O CM CM UI U UI CM C CM CM CM CD CD O O O O O O O - — S — — — — — — CM — CD — — — — — CM CN — — — — — — — — — — LO Ul lΩ CO S S S S S — — CM CM CM CM CM — — CM CM CM CM CM CM CM CM CM —
CO — — —
I I TZ I I I O S CN S l I lΩ l I X l CD OO l I — sf _-ooocooiH_-Icostcoco -HlIIII oo — scoXIII O CO S O 01 X XXII I I I LO co I s
S 00 I CO CO CN CD CN CM CN CD IΩ LO CN CO — CM CD CO CO — CD S l CD CD CD — Ol tD O CO CD CD - I — sf CD CM CN CM CN S U1 S CD CO LΩ S — H U CM CO — O — CM S CD O CO CO CD O S
CM OO 01 O — sf CD LO sf cO st - CO — — — O st CO OO — O — Ol CD sf CM O CO O CD OO CD O S CN - CD CO sf CN CD CD CO S S CO lΩ st st CO CN CD 00 O — CD CD LO sf O CO CD S CD LO 0
— — CD — CD CO S CO CO sf CD CN CVl Ul st CO S S CD Ul CJl © co s co co co oo — co — CD s — oo oo co co oo σi co co sf S O O - — — CO Ul CN CD CM U CO CD S O CD CO st Ul CO sf S U
— O Ul CO CO CO o st co s o s ui oo ooco σi — coco o — — S CO st S OO CO S O CD CO — O CO CD CO CVl sf S S — CO — — CO sf sf Ol S CN O sf CO CD 00 S CO O S S 00 S — CO C CM O CD — OO sf CO CM — O — — sf UJ Ul OO — OI O CO CD OO CO CD CO S O OO CO CO CD CO OO CN — sf oO CO LO OO CN Ul — — CN CM CO CD CD OO LΩ CO CN S — CO Ul O Ol Ul 00 CD CO — CD — 0 S CD S CN CM S sf sf sf sf OO Ul st - — CD U CO — CO Ul — — O — CO sf CO OO CM S - CO CM CO UI CO CD — O CO CO CM CM CM CN OO CO CD CN CM CD — CM CM CD sf Ul CM sf CVJ S CM CO — — C CO CM S f CM LΩ CD CD CD CDCO — COsf sf — Ul CO CDCO — CO COS CD CO CO CO— — CO CO CO CO COO — — CO l CO Ul CO CO CD CO— Sf Sf LO CD— CO CO CD — — CM CM CM — CM — CO COLO C
vo o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o O O O O O o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o so lΩ Ul LΩ Ul Ul Ul Ul Ul Ul Ul LΩ Ul LΩ Ul Ul LΩ Ul LΩ lΩ lΩ LO LΩ lΩ Ul Ul Ul Ul Ul Ul Ul LΩ Ul Ul LΩ LΩ UI U UI UI LO UI UI UI UI UO LΩ UI LΩ UI UI UI UI LΩ UI UI UI LΩ LΩ U LΩ UI LΩ LO U Ul Ul U sf st sf st st st st sf sf st st st st st sr st st st st st sf sf st sf sf st sf sf sf st st st st st st st sf st sf st st st sf st sf st sf st st sf st st sf sf st sf st st sf st st sf st st t sf sf st st st st st sf st sf st st st st st st st sf sf sf st st st st st sf sf st st st sf sf st st sf sf st st st sf st st st sf st st st st sf st sf st st st st sf st st st sf sf sf st sf st sf sf sf sf
© S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S s s s s s s s s s s s s s s s s s s s s s s s s s s s s s s s s s s o CD CD CD CD CD CD OI OI CD OI CD CD CD OI OI CD OI CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD O CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD OI CD OI OI OI OI CD CD CD CD CD CD CD oi oi O) σ) σ)
st st st st st st st sf st st st sf st st st sf sf sf st sf st st st st st sf sf sf st st st sf sf st st st st st sf st st st st st st sf st sf sf st st st sf st st st st st st st sf st st st st CO CO CO CO CO CO CD CΩ CO CO CO CO CO CO CO CO CO CO CD CO CO CO CO CO CO CO CO CO vD CO CO CO CO CO CO CO CO CΩ CO CΩ CO CO CO CO CO CO CO CO CO CO CO CO CO Φ
C C C
o o o o o o o O O O O O o o o o o c> c> o o o o o o o c> o o o o o o o o o o o o o o o o o o o o o φ φ φ φ φ φ φ φ φ φ φ φ υ o φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ cn φ φ φ φ φ φ φ φ φ φ φ φ φ φ
TJ XJ J TJ TJ TJ TJ TJ TJ T X) TJ TJ TJ TJ TJ TJ TJ TJ J XI TJ TJ TJ TJ J XJ TJ Ό Ό Ό Ό Ό Ό Ό Ό Ό TJ Ό Ό Ό Ό
Ul Ul LO Ul Ul Ul U ui i ui LΩ LO LO LΩ LO UI LO LO LO ib ui ui ui ui ui ui ui ui ui ib Lb ui Lb ui ui uiui i ui Lbib Lo
01 CD CD CD CD Ol CD CD CD Ol CD 01 CD CD CD CD CD CD CD 01 CD CD CD σi σ) CD CD oi 01 CD CD CD CD CD CD CD CD CD CD CD Ol Ol CO CO O CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO ro ro ro ro CO CO CO co co co co co co CO CO CO CO CO CO CO CO CO CO CO CO CO CO C CN CM CM CM CM CM CM CN CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM CN CM CM CN C CD CO CO CO CD CD CD CD CO CO CD CO CD CD CO CD CD CD CO CD CO CD CO CO CD CD CO CO CD CD CD CD CO O CO CD CD CD CD CD CD CD C CO CO CO CO CO CO CO CO O CO CO CO CO CO CO CO CO CO CO ro CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO C CO CO CO CO CO CO CO CO O CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO C
st
S CO CN CD — — O O O S O OI CM CD CO CD — CD sf CD LO 00
CM S S sf st S OO CM LO S Ol S O st CM — — O CN O LO O CO — S — CO CO CD OO O CO — CO CO st CN st CM CD OO CO CM sf S CD — O OI CO — S CO CO S S CM sf CM CM CO LO LΩ CO CD S C S CD CO CN st — S S O CD — CO O — S S st CO CD CO CM CO OO CM — O OO — CM — CD CO st cO CO CD CO st CD LO sf CO CO CO Ul CM CO CD lΩ CN — O CO CD O sf CD st CD CM O CD Ul — U1 CO C CO CO CN CO LΩ CN CM LΩ CO IΩ CD CN IΩ CO UI LO CO CO S OO CD CO CD CD — — Ol — — — — — — — — — — — — — — — — CN CN CO — CO CO CO CO CO S — LO OO LΩ LΩ — CD CD — CD CO CD - C
CD — S st CD Ul — CO IΩ LΩ OI S O — 01 CD CO
S CO — UI O CD CD U S OO OO CO CM — Ol sf CM CM S CM S S — — — CO CO OO CO S OO CD CO — CO CD O O O CM sf S S O sf sf st Ul Ul Ul CD - O CD CD sf O CD
— O O — CO st CΩ CΩ S CO CO OO CO CO st CM O CO S O — CO LΩ S O O O IΩ IΩ — O O O O — CO CM CM CO CO CO CO CO CD S S Ul sf O O — CN CO CM CD CD CD CO CO O CO — co o co
S CO — — — — — — — — — — — — CVI CO sf sf st Ul lΩ Ul CD CD OO OO OO OO OO CD — — — — — — — — — — — — — CM U U UI CD — — C CN — CM CM CM CM — CM CO — CO CO sf -
TZ CD CD ∑Z cn CD '— '— — — — O s- "-
Ol I X X I X I H X H g I X CD H I LΩ UI I S LΩ I LL I I LO I I I I X st I I i L lL S l g r l t lD CD l O - I I
— S S O Ul CM s ro CO COI S I CO S Ul Ul X J I I H — CM S CO S CO CO CD LO Ul CM CO CD S LO CD CD LO S U — — O — sf i
Ul CD CD LO O O to ^ o en — 00 CD CO Ul O sf — sf CO CO sf CO CD S OO CO CN CO CD CO CO O S CO CO O CM CM CO S CM CO CM O CO O OO OO IΩ IΩ CO CO CO CM CO CO O C
CM S CO CD sf O ^ Ul 01 S CD O CD O — O CO CO O CO — OO LO sf S — CD O O S — O LΩ sf sf sf CO CM LΩ Ol S LO sf Ul CO CO CM O CO CD OO st Ul CD CD sf S O CO CN Ul o o CD sf O sf 00 CD CD C C CN CVJ CD CN — CO CM sf CD LO O O O CO CD — sf S — — CM Ol CO CO LO IΩ O IΩ IΩ IΩ S LΩ LΩ S CO — — CO S
S CD U U S Ul ui st st st o in s sM sD
CM CM S CN IΩ CN O CO CO O S S S CD 00 sf S S CD CO CD CN CO S CD — CM — CO CO S CO CO O CD CD O S — C
CM — O CM CO CO , — — - o ^ ^ CM , — CD CD CD — 00 — — OO CO CO sf CM CM CO CO — Ul sf 0000 — O CO CM — OI CN CD — — CO CM — CM CD — CM CO —
CD COCO — CO CN σ>_o CO CD CM CM CD CM CM CO CDUl CD COUl — — CM CDCD CO CN CM CM COCO CN CN CDCM CDCM CD CD CO CDCO COCM CD Olsf CO
UI UI U UI UI U LΩ LO UI UI UI UI U IΛ UI UI UI LO UI UI UI U U IΛ LO LΩ LO LO UI U UI U LO UI UI IΩ CD CD CD CO CD CO CD CD CD CD CO CO CD CD CD CD C^ CD CD CD CO CD CD CO CD CD CD CD CD CD CD CD CO CO CD CD CD CD CD CD CO CO CO CD CD CD CD CD CO CD CD CD CD CD CD CD CD CO CD CD CD CD CD CO CD CD CO CO CO CD tD
lO CVI sf CO O O O O Ul Ul lΩ Ul — — — — LΩ CD — CO — — CM CM — — — — C — — — —
ό ssssssssssssssssssssssssssssssssssssssssssssssssss ssssssss ssssssss
CD CD CD CD CO CO CD CD CD CD CD CD CD CD CD CD CO CD CD CD CD CD CD CO CD CO CO CO CO CO CD CD CD CD CD CD CD CO CD CD CD CD CD CD CO CO CD CD CD CD CD CD CD CD sf
_U
X) ca f-i st S CD sf S st OO CM OO — 00 CO CD — Ul — CO CN IO S — CD CD CO CD CD LΩ S Ul sf Ul CVl sf CM Ul Ul — O CM S CVJ
UI CD LΩ U CO CO CD OO LΩ S S O O — CD — sf O lΩ CO LO O O O CO — CD O LO O LO O O LΩ LΩ O LO LO O CO O LO Ol CM CO CD st LO CD CO — S sf CO CM sf - CD CO CO CN OO CΩ LO CD O s
CO CO Ul CO Ul CO Ul CO Ul CO Ul CO CO CD LO CD sf CD Ul st st CO CO lΩ sf sf lΩ CD Ul CO Ul CD CD Ul LΩ CD CD CD CD Ul CO - CO CN CO S CN S CD CO OO LΩ CΩ CD CN S CO OO CO IΩ CΩ LΩ LΩ CO S O
CM CN CN CN tM CM CM fM CM CM CN CN CN CM CN CN CM CN CM CM CN CN CN CN CN CM CN CM CM CM CM CM CM CN CN W
O — CD S OO OO Ul CO CD O CO S sf CD OO S O CN Ul CO CD CD O O O - sf CO OO CD CD Ul O st CD sf CO sf oO st CO
00 0000 01 0) 01 — — CN CO CO st LΩ LΩ CD CD S S S OO OO OO O O O O O — — CN CO sf CO S S Ol S S S CO S O OO CD CO CD LO CO O CO C
O O O O O O — — — — — — — — — — — — — — — — CM CM CM CM CM CM CM CM CM CM CM CVJ CN CM CO CO CO CO st Ul LO OO OO CO CO st OO S OO OO OO Ul S S CJl O — O — CM — — CO CD C
CN CM CM CN CM CM CN CN CN CM CN CN CN CN CN CN CN CN CM CN CM CN CN CN CN CM CN CM CN CN CN CM CN CM CN CN CN CM CN CN CM — — CO CO CM CN CM — — — — CM CO CO CO — CN CM CN S S CD CΩ st CD C
— — CO
I l H rZ X l H l H l σ CO CD CΩ H CO l
" " — -"- — — - - _. OI CD LO U — OI CD LΩ — sf lloiXlII — — CD OO CD S CM CO CD CM I I O l I — I II - I I I I I CD I I Tz oo ∑z I I X X o r: co s co X σi oi CN CN — CO σi OO CD LO S LO CN CO S CO — CO O Ol S CM st S O CO st O S st LO S Ol S — O CD CD CN — O CO OO I CO I 00 CM CM S 00 Ol X st OO CM CO — — CO sf OO — CN — sf — CO Ul CD CO CN lO OO CD CO CD CO CD CD CN st CO — 00 — sf cO CM S — CO — CO O O CO CO — CM 01 sf O — 00 O CD S CD CO 00 O 01 LΩ CO S CD CM C CO — sf st CD CD S S S sf CD S O CD CD sf OO CO O CO OO CN CO OO CO — CO S CO O CD CN LΩ O CD O Ul S LO CO st O CO CO CO CD st S CO 0 Ul CO 01 sf S CO CM O CM 00 S sf U O S S C CO CM — CN — — CO O Ol — sf CD S OO — CD OO S CO CO CD LO S CD S CO CO CO LΩ CD CO CD CO O CD O CM IΩ O CD CO OO CO CD CD IΩ CD Ul CD S S CO Ul Ul CO CO CD 01 CO f sf CO Or- Or- — O CO LΩ CN CM CN LD — LO OO CO CN sf CM CO OO CO CD sf CM O CD CJl CD OO sf LO st OO OO S st cD lO — CM sf CD CO LΩ CM CD st oO CM Ol LΩ S 01 Ol Ul CD CO 01 CD CD st CD O co i s o oi O IΩ OO CD OO OO CD CO S CO — CO sf st cO Ul O st Ul sf — CO LO CO Ul Ol sf CM st - CO st st sf sf cO O CD CN S CM S - CD CO CO S — CD S 00 sf CM — CO 00 sf — O — CD co σi oi C σ o CO CO CN CD — — — sf — CM CO CO CO CO— COsf CD COCM CM CDCM sf CM Ul Ol CD CD Ol CD CD CD CO CO CDsf CO CO CO CDsf CDCO CDCO CO CDCO CO — CM CO CDCO — CD CDOO CDUl sf U) CO CDS C o o o o o o o o o o o o o o o o o o c> o o o o O O O o o o o o o o o o o o o o o o vo φ φ φ φ φ φ φ φ φ φ cn cn φ φ φ φ φ φ φ Φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ o XJ TJ TJ TJ TJ TJ XJ XJ XJ J XJ TJ TJ TJ J TJ TJ TJ TJ TJ TJ TJ XJ XI XJ J XJ TJ TJ TJ TJ TJ TJ J so ui ui ui ui ui ui Ul Ul Ul U u ui ib ui ui sf" sf sf sf sf sf sf sf sf sf sf sf st sf sj sf sf sf sf s
O) CD O) CD CD CD CD CD CD CD CD CD CD oi oi OOO O O o o o o o
CO CO CO CO CO CO CO CO CO CO CO CO CO co co d CD dCD CD CO CD CD CO o CDdCDdCD d CD dCD o CDdCDdCD o CDdCDdCD d CD C
CM CM CM CM CM CM CM CM CM CM CM CM CM CM CM CD CD CD CO CD CD CD CD CD CD CD CD CO CO CD CO CD CD CD C
© CD CD CD CD CO CD CD CD CD CD CD CD CD CD CD LO LO Ul LO Ul Ul Ul Ul Ul U Ul Ul U) Ul Ul Ul LO LO Ul o CO CO CO CO CO CO CO CO CO CO ro CO CO CO ro CO CO ro
CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CM CM CM CM CM CM CM CM CM CM CN CM CM CM CM CM CM CM CM CM CM CM CM CM CM C
CD CO CD CD CD CO CD CD CD CD CD CD CD CD CD CO CD CD CD CD CD CD CD CD CO CO CO CD CD CD CD CD CD CD CD CO CD CD CD CD CD S S S S S S S S S S S S S S S S S S S S S S S S S CD CO CO CD CD CD CD CD CD CD CO vD CO CO CΩ CD CD CO CO CO CD CD CD CD CO CO CD CO CD CO CD CO CΩ CD CO CO CD CO CO CD CD CD CD ω
CM — C
CD C S
I
cvi Cvi cvi d d d s s s oo
_ X)
CM CO — CO S O CO S CO S S CO S O CO O CD CO CO CO CO — CO CM CO CM O LO σi st O CN CM CM O CO O OO st CD S lO CD — CO st CM sf CO CN lO - — CM CO — OO sf CO LΩ LΩ CN CO OO CN lΩ C O O — CM LO — sf O O O O S CO — O Ul CD LO O sf O O O CD O O O O O sf S — Ul — — O CD O CD O — CO CD — O CD — O O — CO — O — O CD CD O O CO — O O O O O lΩ lΩ lΩ CM CO LΩ sf LΩ lΩ Ul Ul CM CM LΩ LO — st sf lΩ — Ul LO Ul st lO LΩ CN Ul lO CM CM CO CN CO CO lΩ CN Ul CM Ul CO st — LΩ Ul CO Ul LΩ lΩ Ul st Ul lΩ Ul Ul sf sf Ul lΩ CO Ul Ul lΩ Ul lΩ lΩ st CD CM CD Ol — CO CO CD CO CΩ Ol sf Ol Ol CN LO Ul CD CN sf CD CD CD lΩ CD O lD CD OO CO S S CD CD sf cO CD S CD — CD OO O CM O CO OO CO CD CN CD LΩ CD O CO CO C
CM CO LΩ CO st st lΩ Ul LΩ LΩ CO CN CM CO CΩ CM S S S CN O O O O — CN CO CM CO CD O sf — CM CM CO CO CO CO sf st st st LO S CD S OO OO CD CD O st st UO Ul Ul S S CD CO Ul Ul S S O
— — — LO — — — — — — — CO CO — — CD — — — CD CN CN CM OM CM CM S CM CN S CD CD O O O O O O O O O O O O O O O O O O O — CM CN CM CM CM CM CM CVl CO CO CO CO CO st s
— — — — — — — — — σ) θ) - — σ) — — — O) - — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — —
— CM — — CD '- CD — — — — — — — — CD ^
CO 00 I I U I 000000 — I I — IΩ I H I X I O LO CD CO O I I LO - S I Ξ I loIcor:olooιzcMor:scDoocNHcocD- co — co co co CD r: — ists
O CM O CM 00 st CD sf LΩ sf sf CD LΩ LO O S CN CM CM st LO O CO OO OO O CO sf O CO CN CO l — Ol Ul CD sf CM CD CD l oO — I st O st st cO CO CD CD Ul Ol CD CM O S O — I sf — LΩ
O sf — CO LΩ 10 S 00 CO LO CD Ul CO S CΩ LO lΩ LΩ LΩ CD CN CN CD CO OO S CD OO CO OO LΩ st S CD O OO S S OO OO OO CD S CO Ol sf S sf Ul CM O — CO st sf S CD αO CD S CM CO sf OO Ul CO C
O CD CD 0000 O S 00 sf CO sf sf CN 00 — CO CO CO CO CO O IΩ CO CO OO O — st CD CO — sf oO CO CN O O sf — IO — S CO U S CO — CD CN S S S S CD — CO CD CD O CD — OO sf lΩ O sf C
CO O LΩ S CO — CO S LO sf 00 Ul CD O U O CM CN CM O U CN CN CN S — CD LO OO OO CD S CD CO sf OO CD CM sf - — CO LO O CD OO O S — S S S LO CO CD st CD CD Ul CO S S S CM O O O
CN St CM O LO Ul — S 00 — 00 sf S CM O S CO CO CO S CM — CD sf CM CM CM — O CM sf O CO O S CO CO CO — O CO CO OO O CO OO sf S — — LO O LO CO sf CD CM CD — S CM S OO LΩ OO S
CO — 00 sf sf sf 00 CM CM CO CO 00 CD CM CO LΩ — — — lΩ CO — — — CM LO sf cO sf CD — OI C CD CM CM OO CM st CO LΩ — — f - O CO CN IΩ LΩ — st C CN — OO S CN CO CM — CO O sf CD CN
CD CO COCM CO COCO CO CO CO COCO CVJ CO CO— CM CM CM — CD CD CD CD CDCO CM CO CO CD CDLΩ IΩ CO CN COCN COCO COCN CD— C0 C0O CO CO CO COCM C0 CD CD CD CD CD CD CD C0 CD C0CO CO— CO o o o o O O O O O φ oφ φ Φ Φ Φ Φ Φ oφ oφ oφoφo O O o O O φ o φ φ CD φ oφ oφ oφ Cύ oφ o o o c> o o o o vo φoφ Φ Φ α> oφ oφo φ φ φ φ φ oφoφ oφ o u TJ TJ J TJ TJ TJ TJ TJ TJ TJ TJ XJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ J TJ XJ TJ TJ TJ TJ TJ J J so sf" sf" sf" sf sf st sf sf sf' sf sf sf sf Sf st" sf st sf" sf st sf f sf st sf sf sf st sf sf sf sf sf sf st sf' sf sf o o od d d O O O o o o o o o o od o o o o o o o o o o o
CO CDdCD co co d CD dCD d CD dCD CD CO CD d CD dCD o CDdCDdCD d CD dCD to o CDdCDdCD d CD dCD d CD dCD o CDdCOdCO CD CD CD to to to d CD dCO CoO CdDdCO to CO CD CD co co CD CD CD CD CD CD CD CO CD CD CD CD CD CD to CD CD CD CD CO CD CD CO CD CD CD CD CD to to CD CD CO CD CO
© U) Ul Ul ui ui Ul LΩ LΩ LO LO Ul Ul Ul Ul Ul Ul LΩ Ul Ul Ul l Ul Ul Ul LO LO LO Ul Ul Ul Ul LO u> u> U) l U) Ul U o CN CM CN CM CM CM CM CM CN CM CN CN CM CN CM CM CN CN CM CM CM CN CM CM CN CM CM CN CN CM CM CM CM CM CM CN CM CM CM CN CM W ssssssssssssssssssssssssssssssssssssssssssssssssssssssssssssssssss
CΩ CΩ CO CΩ CO CO CD CD CD CΩ CO CO vD CD CO CD CD CD CO CD CO CD CO CD CD CD CΩ CD CO CD CO CO CO CO CD CD CO CO CO CO CD CO CΩ CO ω
CM CO CO C CO CM S S CN S S CO st LO S CN OO CO S CO CO CN CO — — — OO sf st CO CO sf CN O CO CD OO CD CM CO CD sf CN S CD CD CO CO CO st CO CD CD - — CO sf lO - OO Ul Ul CO sf CD sf sf CN S sf CO CD — CO CO OI CO CD LO — CD CO — S CO O CD O — O OO S O OO CM O CO CN S CO S S CD OO — CM O — st st Ul — CD CD LΩ C OO CD OO S — OO LO — st CD CD O — CO — CN CM CO CO CO CN sf sf st Ul Ul Ul CO CO OO CO CD CO S S OO CD OO Ol - — O CD — O CM — — — O — — — — — CM CN CM — — CO LO CN — LΩ sf OO IO LΩ LO LO CO OI — OO CD OO — — — — — — — — — — — — — — — — — — — — — — — — — — — — CM CM CVJ — CM CN CM CM CM CM CM CM CM CM CN CVJ CM CN CN CO CO CM CVJ CO st CO CN C vo
IΛ CO CD sf CO CO O CO st CO CM Ol - LO LΩ LΩ CO CO CD OO CO CD OO CM O CO S CN CM CM UI U S S - — CO O - — CD JN O O S O CD O CD O — OO CN CD S CO CD — OO CD CD CD OO CD CD LO Ul CD CN CO S CD CD CD st CD CO sf — CO CO LO OO CD CD CD CO S S S S CO CO CD — — CN LO OO OO
© Ul LΩ S O sf — sf OO sf CO O O O st oO CD CD S O O O — — CN CM CM trj C CO CO C C st st Ul LO CD CD CD S S S S S 0000 C0 C» CO t» CO t» CD σ) CD CD CD CD LO CM Ul st CM CO sf 00 0 © CO CO CO sf sf Ul Ul CD S S OO CO OO OO OO CD CD CD - — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — UI UI UI CD CO CD OO OO C Ul
H — — — CM — — — — — CD '- co ■■- s- >- U r-rriri-iuiii I L_ I III I I X O X I CO I I I I I ΣZ II I I Jg ∑Z I I I I TZ I I I I I I I I sf I CD I I XXcDlIs scooo CD CD S st LΩ — co co O — — CO — Ul sf Ul CO O CD CO CM sf sf CM S CO — O O I sf — S CO sf CM X I sf cO - Ol CD l ul O CO — S sf Ul OO CO CM — s —
PH sf sf — — CO — Ol CM sf CO sf cO S OO sf — — ssss CM — O CO S 00 CO O CD 0000 Ol O CO — CD CD sf O O CO O S CO Ul LO CO CO OO — CD S CO CD S CO sf S OO CO sf CN CN S CD CD CD S — CO IΩ OO IΩ LO OO CO CN C IΩ CD — CD CM CD CD CN — CD CO 00 — 00 CM CD CM O O 00 — sf S CD — CD O CO CO O CM sf CO CO — OO S CO CO OO CD O O sf CM st CO CD O — CD — S CO CO CD IΩ OI UI UI — O CD OO — — Ul sf CD Ul sf CD CD O) CO — CM LO CO CO — LO Ul S S Ul CO o CO CM CO LO CO CD CO CO sf S CO CO sf cO CD S Ol — σi CM CM CO CM O O CΩ CO CO CO CD CO C σ) σ) σ) st σ) θ co co cM oo cM sf ui o o σi o S CD CD CO O CO — S CO CO CD O sf S S 00 s s S O CD 00 to O S 00 Ol LO to CO S S C0 0O CO S — LΩ sf — CO S OO LΩ lΩ CD CD OO S s
0O 0O LΩ — CD CD — sf O st — CO CO LO O O CM S O O CD CO CD CO S sf O O S LO O O — sf — CD CD CM 00 CD — Ul CD sf CO CD CO CD CM CO CD S Ul OO CD S CO sf O LO sf st cO — O CD L S S CO — LΩ U CDCD sf CN CO CO CD CO LΩ st CO CDCM CM CO CN CM LO OICO CO Ul CO CDCN CM COCO CM CO CO LO S COCD CO COCM CO CO CO sf CM CM CM COCD st CO Ul — CM CO CO CO COCN COCO CO
O O O O O O φ φ φ o o o υ o o o o o O O O O O O O O o O O O o o O O O O O O O o O O O O o O o o O O O o O O O O O O O O O O O O O O O o φ Φ Φ Φ φ φ Φ Φ φ φ φ φ φ Φ Φ Φ Φ Φ φ φ φ φ Φ Φ Φ φ φ Φ Φ Φ φ Φ Φ Φ φ Φ φ φ φ φ Φ φ φ φ Φ Φ φ Φ Φ Φ Φ Φ φ φ φ Φ Φ φ φ φ φ φ 0) TJ TJ TJ TJ TJ TJ TJ XJ XJ XI TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ X) XJ XJ XJ XJ XI TJ TJ XJ XJ X) XJ XJ TJ TJ TJ TJ TJ TJ TJ J TJ TJ TJ TJ TJ XJ TJ XI TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ Tl J TJ TJ CM CN CM CM CM CN CM CM CM CN CM CM CN CM CM CM CM CM CM CM CM CM CM CM CVJ CM CN CN CM CM CM CM CM CM CM CM CM CM CM CN CM CN CN CM CM CM CM CM CM CM CN CN CN CM CM CM CO CO CO CO CO CO CO CO C
OOO O O O O o o O O O O O O O O O O O O OOO O O O O O O 0 O O O OO O O O OOO O O O O OOO O O O O OOO CN CM CN CN CN CVJ CM CN C sf sf sf sf st st sf st t st sf sf sf st st sf st st st f sf t sf sf st st f st st sf st sf sf sf sf st st sf sf st t sf sf st f sf st sf sf osf ost sf sf sf sf sf sf sf O O O O OOO CD O) CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD Ol 01 CD CD CD O) CD O) CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CO CO CO CO CO CO CO P0 C
00 00 CO CO 00 CO 00 CO C
CD tD CD CD CD CD Ol CD CD CD CD CD Ol CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD Ol Ol CD CD CD CD CD CD CD CD CD CD CD CD CD CD Ol CD CD CD CD CD CD CD s s s s S S S s CO CO CO CO CO CO CO CO CO CO CO CO O CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO O CO CO CO CO CO CO CO CO CO CO CO CD CD CD CD 01 CD CD CD C co co co co co co ∞ oo rø co co m ∞ oo co oo co αD oo co co co co co co rø
CO CO CΩ CΩ CD CD CO CO CD CO CD CO CD O CD CD CD CD CO CD CO CO CD CD CD CD CO CD CD CO CO CO CD CD CD CD CD CD CO CO CD CD CD Φ t
_2
X)
C3
H
UI CO CD CO CD CD CD — CD — CO O CD CO CO S CO CD CO CN CO O O — CO O CO LO OO Ul Ul S st OO O Ul CD Ul S Ul lO LO CO st O Ul CD LΩ LO O lΩ CD CD CO st Ul Ul LΩ CO - L O OO CD CD sf Ul Ul O CN CM CN Ul CM OO sf CD Ul CO CM CM CM st CD CD - — Ul Ul Ul sf CD LO — LO CD LΩ IΩ O UI — O lΩ sf Ul — lΩ LΩ lΩ Ul st S O CN LO LΩ LO LΩ CN — CV1 CD LO LO O CN — CM S lΩ CO CD CD CM CO crl CO CO CN CO CM CO CM CM CO CO CO CO CO OO OO CD CM CM CM CM O CM CM CM CM - CM CM CM CM CN CM CM CM CM CM CM CM CM CM — — CO CO CM CM CM CM CO — CO — CN CM CO CO — C st ui to co cD co co co co co co co co co co co crj co co co tri co cM CM CM co co co co co co co co co co co
CO CN S OO st CM O O CO lΩ CN sf S — S S CO CO OO O Ol CD CO CO CD CO sf CD OO Ul CD O OO CJI CD CD CD CD CD O O O O O O CD st O CD IO O CD O OO O O O CO CD O s CO S S OO CO O O CM CO st cO sf lO Ul Ul Ul CD CD - CN CO sf — sf CO — — S OO CD O LO OO OO OI O O — CN CM CM CM CM CN CN CO CO CO CO CO CO UI CD CO CD CD CO S CO S CO CO CO S OO CO C LΩ Ol CN CM sf Ul CM CM CM CM O O O O O O O O — — — — S S S OO CO CO OO OO OI CD CD CD CD O O O O O O O O O O O O O O O O - — O — — O — O — O O O — — O — CM CO CO CO C CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CM CM CM CM CM tM CM CM CM CM CM CM CM CO O CO CO CO CO CO CO W CO CO C^
— CD CO CD — — — — — — — — CD CO CM — CD CM CD — — —
I H H H <£ rr CO I I I co I I I O - _- I S X S TZ I CD CD CO LO — H CΩ H l co I H S O O LO I I I I I — rr oo TZ CO CO — S — X X Ol st CO Ol CM - LO OM - σi oo X cvj ui oo cM l OO H LO CO CD Ul CN CM sf CD I I I
S CD Ul CO LΩ O — CM — S — CO CO sf Ul sf — CO 00 sf i oi x I CM I OOxsf OsOsstoOoiOOiCDisf r st oO CD CO S sf CO O — CN CN CD CD sf CO O S Ol CO — Ol sf Ol O CD sf CO CO LO O CN CO OO CM CO CO S sf CM S CO CD CD CO — CD CM O O CO CD S CO U — CD CN sf — S CO CO CO LΩ — S O LO OO CO CN CD S OI CO S CN — CO CD CN OO sf S CD Ol OO CD — CN O CO CD CO CN CO CO LΩ O sf LO Ul S — — CO CO CD LΩ — CO O CO — CO CD CM O Ul 0 00 CD sf LΩ CM CM sf CD LΩ CN LO OO C CVJ S CM CO — CO O lΩ CO CM CN CD S S sf OO OO S CO O LO sf CO lΩ CM LΩ CO sf O CD CO CO S 01 CD CD CD — O σi st CM O OO S Ul Ol sf lΩ sf — S O 010000 CO CM o o σi co cD O U co CN — OO st CO Ol tD Ol st — CO O CM CD CN CD CD st CD — CO S O CD CM OO CO CM O CO CN LΩ S CO — CO CD O σi CO CO CO LΩ S — CO CO CN IΩ CM sf CO 00 CD CO CD O 00 IΩ CN CD CD CD CN CD CD 0 CO CO S CO S OO CD CO LO CO sf — sf CD CO CN CM CD CO O CN OO — O sf oO CO — CN CD CO sf O — LO sf S — sf O LΩ st CD CD CO st — CVJ CM CO CO CD O LO CO CO sf CM CO — — Ul — — sf S s Ul COsf sf sf S - CDCM — CD CDCΩ CM CM CD CD CDCD CO COCO COCM sf S C CD CO O) CO CO CN CDCO CD COCO CDCM CM OICO OICM CO CD CO CO CD CO CN sf CD CO CO COCD CD CO CO CD CDCO LΩ CO o o o o O O O O O o O O o O o o o o o o o o O O O O O o o o o o O O O O O o o o o o o o o o o O o o o O O O O O O o O O O φ Φ Φ Φ φ φ φ φ Φ Φ Φ Φ Φ φ φ φ φ φ φ φ Φ Φ φ φ φ φ φ φ φ φ o O O vo φ Φ φ φ φ Φ Φ Φ Φ Φ Φ Φ φ φ φ φ φ φ φ φ φ φ Φ φ φ φ Φ φ Φ Φ Φ Φ φ Φ φ φ o TJ TJ TJ XJ TJ TJ TJ TJ TJ TJ XJ TJ TJ XJ X) TJ TJ TJ TJ XJ XJ XJ XJ TJ TJ TJ TJ TJ TJ TJ TJ XJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ XJ X) TJ TJ TJ XJ XI XJ TJ TJ TJ TJ TJ TJ XJ TJ TJ TJ TJ TJ TJ TJ TJ T so CN CN CN CM CN CM CM CM CM CM CM CM CN CN CN CVJ C CN CN CM CM CM CN CM CM CN CM CM CM CM CN CM CN CM CM CN CN CN CM CM CM CM CM CM CM CM CM CM CM CM CN CN CVJ CM CM CM CN CM CM CM CM CM CN CM CM CM C
O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O OOO OOO O O O O O OO O OOO O O OO sf sf sf st sf sf sf sf sf sf sf st sf sf ost of f sf f sf sf sf sf osf sf sf sf sf sf st sf sf o o sf sf sf sf sf sf sf st sf sf sf sf sf sf sf st sf sf sf sf osf sf sf sf Sf sf st sf sf sf sf sf sf sf s CD CD CD CD CD CD CD CD CD CD CD CD CD σ> CD σ> CD Ol CD CD CD CD σ> CD CD CD Ol Ol Ol CD Ol Ol CD Ol CD CD CD CD CD CD CD CD CD CD Ol CD Ol CD O) CD Ol CD Ol CD CD Ol CD CD CD Ol Ol Ol CD CD CD CD C
© o CD CD CD CD CD CD CD Ol CD CD σ) CD CD CD CD CD CD CD CD Ol CD CD CD CD CD CD CD CD CD CD σ) CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD Ol CD O) CD CD CD CD CD CD CD CD CD CD CD CD CD CD Ol CD CD C CO O O CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO CO O CO CO CO CO CO CO CO CO CO CO CO CO CO O CO CO CO CO CO CO CO CO CO CO CO CO
CO CD CD CD CD CO CD CO CO CO CD CD CD CD CD CO CO CO CO CD CD CD CD CD CD CD CO CD CO CD CD CD CD CD CO CD CO CD tD CD CD CD CD CD CD CD CO CD CD tD CD CD CD CD CO
Table 4
978302.3.dec 4049703H1 97 382 69 978302.3.dec 5504612H1 41 112
978302.3.dec 3273789H1 110 357 69 978302.3.dec 1903639H1 42 296
978302.3.dec 4803070H1 311 568 69 978302.3.dec 3580385F6 52 446
978302.3.dθC 5502955R6 345 751 69 978302.3.dec 5537637H1 51 229
978302.3.dec g3330490 449 844 69 978302.3.dec 904267R2 1840 2291
978302.3.dec g916384 496 716 69 978302.3.dec 004607H1 1842 2021
978302.3.dec 4570560H1 508 772 69 978302.3.dec g2141685 1866 2354
978302.3.dec 4799067H1 536 814 69 978302.3.dec g2111916 1877 2317
978302.3.dec 5990884H1 554 849 69 978302.3.dec 5314445H1 1895 2139
978302.3.dec 6421386H1 684 1179 69 978302.3.dec 575430T6 1938 2513
978302.3.dec 449422H1 709 952 69 978302.3.dec 2962839H1 1947 2232
978302.3.dec 1303054H1 746 977 69 978302.3.dec 3580385T6 1949 2513
978302.3.dec 3494973H1 826 1129 69 978302.3.dec 469485F1 1956 2552
978302.3.dec 2762433H1 883 1137 69 978302.3.dec 629855H1 1972 2105
978302.3.dec 671218H1 883 1139 69 978302.3.dec 6590117H1 1976 2552
978302.3.dec 5374645H1 943 1204 69 978302.3.dec 6590017H1 1976 2505
978302.3.dec 4293786H1 962 1115 69 978302.3.dec 4508845H1 2014 2296
978302.3.dec g30733 1008 1310 69 978302.3.dec 2878353H1 2019 2175
978302.3.dec 1556785H1 1045 1237 69 978302.3.dec 2573374H1 2027 2280
978302.3.dec 1557004H1 1045 1247 69 978302.3.dec 2573245H1 2027 2266
978302.3.dec 1554631 H1 1045 1214 69 978302.3.dec 412698H1 2043 2251
978302.3.dec g4220895 1067 2218 69 978302.3.dec g4302126 2084 2552
978302.3.dec 2667190H1 1096 1343 69 978302.3.dec g3884354 2097 2554
978302.3.dec 3389771 H1 1115 1406 69 978302.3.dec g3434749 2099 2552
978302.3.dec g1296252 1145 1477 69 978302.3.dec g5100695 2104 2551
978302.3.dec 495467H1 1144 1392 69 978302.3.dec 4914418H1 2117 2363
978302.3.dec 495467R6 1144 1621 69 978302.3.dec 816019R1 2128 2548
978302.3.dec 469485R1 1184 1693 69 978302.3.dec 816019R6 2128 2481
978302.3.dec 469485H1 1184 1418 69 978302.3.dec 816019H1 2128 2363
978302.3.dec g2783215 1290 1698 69 978302.3.dec g2727063 2138 2488
978302.3.dec 5117988H1 1427 1697 69 978302.3.dec g5631519 2140 2552
978302.3.dec 2656968H1 1447 1679 69 978302.3.dec g4268066 2143 2552
978302.3.dec 044893H1 1457 1621 69 978302.3.dec g3038701 2146 2554
978302.3.dec g2023688 1493 1815 69 978302.3.dec g2905509 2168 2552
978302.3.dec 2649157H1 1493 1722 69 978302.3.dec g2337320 2170 2553
978302.3.dec 5392455H1 1535 1805 69 978302.3.dec g5392280 2174 2554
978302.3.dec 2972663H2 1540 1812 69 978302.3.dec g2433438 2176 2551
978302.3.dec 4887906H1 1548 1784 69 978302.3.dec g4764364 2181 2551
978302.3.dec 1494628H1 1565 1777 69 978302.3.dec g4900354 2181 2560
978302.3.dec 4123045H2 1577 1833 69 978302.3.dec 3455054H1 2195 2459
978302.3.dec 598817H1 1592 1744 69 978302.3.dec 3406692H1 2206 2456
978302.3.dec g1697464 1592 1873 69 978302.3.dec g2767388 2211 2558
978302.3.dec 2782445H1 1595 1839 69 978302.3.dec g3960728 2213 2552
978302.3.dec 3071188H1 1599 1864 69 978302.3.dec 1627118T6 2212 2512
978302.3.dec 1499678F6 1611 1983 69 978302.3.dec g2751958 2213 2543
978302.3.dec 1499678H1 1611 1833 69 978302.3.dec g2969710 2220 2555
978302.3.dec 3605819H1 1627 1786 69 978302.3.dec g2839836 2219 2552
978302.3.dec 2603477H1 1645 1890 69 978302.3.dec 1499678T6 2223 2506
978302.3.dec 5396451 H1 1714 1966 69 978302.3.dec 2244852T6 2227 2510
978302.3.dec g1925214 1718 2183 69 978302.3.dec g2849258 2234 2552
978302.3.dec 6410645H1 1726 2080 69 978302.3.dec 2780422H1 2241 2476
978302.3.dec g1923462 1725 2117 69 978302.3.dec 5037261 H1 2250 2433
978302.3.dec 5563939H1 1739 1949 69 978302.3.dec g1924159 2256 2552
978302.3.dec 3901151 H1 1746 2017 69 978302.3.dec g5657243 2257 2552
978302.3.dec 3467260H1 1765 1865 69 978302.3.dec g2111854 2259 2559
978302.3.dec 2153778H1 1772 1913 69 978302.3.dec g1925215 2269 2563
978302.3.dec 2363625F6 1792 2120 69 978302.3.dec g824152 2307 2563
978302.3.dec 2363625H1 1792 2020 69 978302.3.dec 2511077H1 2321 2549
978302.3.dec 6167007H1 1803 2363 69 978302.3.dec g1265044 2322 2552
978302.3.dec 2363625T6 1814 2354 69 978302.3.dec 627673H1 2324 2549
978302.3.dec 043817H1 1819 2062 69 978302.3.dec g3118123 2331 2552
978302.3.dec 1627118H1 1835 1935 69 978302.3.dec 927354R1 2346 2599
978302.3.dec 1627118F6 1835 2274 69 978302.3.dec 927354H1 2348 2593
978302.3.dec 904267H1 1840 2062 69 978302.3.dec 3241557T6 2366 2510
978302.3.dec 4921037H1 1 290 69 978302.3.dec 790882R1 2367 2552
978302.3.dec 4913159H1 31 294 69 978302.3.dec 790882F1 2367 2552
978302.3.dec g4838128 44 2568 69 978302.3.dec 039176H1 2375 2542 LΩ st st OO CO CO CO CO CD CO CO LΩ lΩ LΩ CD — CD LO sf — Ol OO OO OO OO O Ul — CD CO O CN CO S IO O LΩ O O IΩ S — sf CO CD — CO CO CD CN UI CN O CO CM LΩ U S OO CD
U1 O O 00 — — Ul CO CO OO CN OO OO OO CM CM CD Ul Ul S O sf Ul OO st - S CO CM LO sf CM r- O CO CD OO OO CD OO S sf CO CD CO LO Ul Ol CD — sf oO Ul sf CN CO S S CD CD CD Ul st CO Ul CM s
CD CD CD CD OO S S S S CO CD O O O S CD lO CO CD Ul Ul CO C» S CD CD CD S S CD S CD CD CD O) CO O CO O t» OO OO CD CO C» C» CD S S t» CO CD CD CD CD CD CD CD
— — — — — — — — — — — CM CM CN — — — — — — — — — — — — — — — — — — — — — — CM — CM — — — — — — — — — — — — — — — — — — — — — sf sf Ul Ul CO CO C vo
IΛ S OO OO S CO S CM CM CO — CM sf cO CN CM OO O O CO st CO — CD OO CD OO st s CO CN OO sf — — CO CO CO CM CD O — CO CO OO OO CD O LO Ul Ul Ul S CD O CO OO CM CD S st JN CO CO CO CO CN Ul CD CD S CO S CD CD OO LΩ S OO OO OO CM OO CD CO st CO CO st st — CM — LO CO CD CO CD S — CO sf OO OO CO CD CD st LO Ul S S OO OO O CD Ol CD CD CO — CO CM CM CD — — sf L
© sf sf st sf lΩ st sf st st CM C CO CO CO st cO W CO CO CO CN CO Ul Ul st st st st Ul Ul Ul CD CD CD CD CD CD CD CD CD lO LΩ LΩ lΩ Ul lΩ LΩ LO LO U n © — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — CM CM CM CN CO CO C Ul
H U 111 TZ TZ oo I co 111111 rr 1 rr rr 111 - I rr 11 tiioi iinirirr I I I I I I ) —. CN 00 I I I I LO I
I I I 00 CD CD CD CN st l l cD Ol S st st O O tD S S l CD S l I lO sf sf CD OO l Ul sf Ul CM sf CO CD CD l O OO CM CN CO l oO l I CO sf CD st CO sf CM U sf CO Ol sf st S CD I S — CD IO OO CN st — CM CM CM CO — CD CO CD CO sf S - CM CO OO O CD OO CD CD CD — CD st lΩ Ul st lΩ CM O st CO Ul CO CD CD CD CVJ CO CO CD OO O CM O sf CO CN sf CM sf CO CO — LΩ S CM Ul Ul S O 00 sf O lΩ st O O lΩ CO CN — CO CO CN st CN O S CO CO CO st O OO LO — Ul OO S CO S Ul CD CD CO CN OO CD S CD CD st CM O CD U) O CD CO O s CD CD LO ) s CD CD CD CO CD S C S sf sf CD CO — sf - tD S S CD OO OO CM S CD CD CD Ul CO O OO lΩ CN OO lΩ CM O CO O CO CM sf sf CO OO O CO CO S CO - — sf CO LO O LO CO st S CD O st CN LO LO o CN CN CD CO CM 00 C CM CM CM CO CN LΩ sf S CD Ul CM CO CO S CO CO CD CVl S OO Ul LO CM CM - — OO O O CO O CD LO IΩ CO CO CD CO — CO sf CD LO CM O O — LO — CD CO CO CM CN CD CD CO sf C — — — CD CO S CO S O CN CO U O CM CD UJ CM - σi Ul S Ul Ul CM O CO — LO CO — OO S S S OO OO st S sf CD O lO CD CO O — CM sf sf - ,— CM CD s LΩ O CM o
CD sf s S CO CO CN sf sf sf 00 S — CM L CO CO CO CM CM CM st CO CM O 00 COCO COCO CD — CO — CN CDCD CM — CO CM LO CO — CJ>— sf — — CD CO COsf COsf O st Ul LO lΩ CO CM CO LO OO CM LO CM r— O CD CD Ol Olsf sf CO — CDCO
O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O
Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ O O O O O O O O O O TJ TJ TJ TJ TJ XJ XJ XJ XJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ T^ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ
— — —— —— —— —— —— —— —— —— —— —— —— —— —— —— —— —— —— —— —— —— —— —— —— —— ——— —— — —— —— —— ——— — — — — — — — — — — — — — — — - — —'—- — — — — — — -q-q-q-q-q-q-q-q-q-q-
CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD O) 0) CD CD O) 0) CD CD O) CD CD O) 0) 0) 0) 0) 0) 0) 0) 0) 0) CD CD CD CD O) CD CD CD CD CD CD CD CD CD CD - — — — — — — — — —
IM CM tM CM tM CM CM CN CM CM CM CM CM tM CM tM CN CM tM CM CM CM CM CM CN CM tM CN CM CN CM CM CN CM CN CN CN CM CN CM CN CM CO CO CD CD CO CO CO CO CO CO CD CO CO CO CO CD CO CO CD CD CO CO CO CO CO CD CO CO CD CD CO CO CO CD CO CO CO CO CO CO CD CO OO OO OO CO OO OO OO CO OO OO OO OO CO CO OO OO CO OO CO CO OO CO CO CO OO OO CO OO OO OO CO OO OO OO OO OO OO CO CO OO OO OO CO OO OO OO OO OO CO OO OO CO OO OO OO — — — — — — — — — —
CM CM CM CM CN CM CN CM CN CM CM CM CM CM CN CN tN CN CM CM CM CM CM CM eM CM tM eM CM CM CM CN CN CM CM eM CM tM CM CM CN CN CN CN CN CM CN CM CN CM CM CN CN CN CN CN tM CN CN CN CM CN CM CM CN CM tM CN CN tM CN CN CN CN CN CN CN W oooooooooooooooooooooooooooooooooooooooooooooooooooooooo — — — — — — — — — — sssssssssssssssssssssssssssssssss sssssssss ssssssssssssssssssssssss
sf Ul CD CM CD CD CD — sf CO CD O O CD st CO O CD CO Ul CD Ul st CO S LO CO OO CN OO CO O CO — S Ul S O sf CD LΩ CD O CO LΩ OO O Ul LΩ — — CO C
Ul OO CD Ul O Ul OO Ol O Ul Ul CO O sf S st O CD OO OO CO OO CO Ol CO S CD OO O OO O OO st CM - CO CO CN st O CO CD CN CO lΩ CO CO O lΩ CO S CD S sf CD LO CD CD CD CO CO OO OO Ul O sf o lO LO lΩ LΩ CD O O O — CO CO sf st S CO σi — O O O O O O O OO — CO S CO O CD CO S CD — LΩ CO O CO CD O O — O — sf - CM — CM CO sf CN CM CM CM CM CM CO CD O O O CD S CD C CN CN C CM CM CM OM CN CM — — — — — — — CM CN CN CM CN CN CN CN — CM CM sf CM CM Sf CO st Ul OO OO CO CD Ol CD - — — — — — — — — — — — — — — — — — — — CM CM C — — —
CO sf st sf — sf Ul S CD CD CD CO CM S Ul OO CM CD O CD O O LΩ CO sf — — CD CD — CO CN CM sf O sf CD CD lΩ - CM — sf Ul CO S O O O CD - — CD CN CD S UI CD S — CO sf sf CO sf Ul CD CD CM S O S LΩ CD LΩ CO st S S O — CO — Ol sf O O O O CO CO LΩ LΩ — OO OO CM CM — CVI CO C st sf sf st st S S S CO - — — — — — S S S CO OO OO OO S S S S S S O CO CM sf CM S CD O CM lΩ CO CD O CN CM LΩ CD CD lΩ Ul S O O O O O O O O - sf sf tD tD st st sf CM CM CM CM CM — — — — — — — — — — — — — — — — — — — — — — CN LΩ CO — CN CO Ul Ul CD CD CO S S OO CO OO OO OO OO CD σ CD — — — — — — — — — — — — — — — —
!— — CN I CD
CM I — ui cD sf uirr l l I X I — — O s" rr r oi CM rz I lull Irrini I I I I 1 r ii 11111111111 s H rr rr
LO sf S CO CM OO LO CN l CO CM to S s I CD I LΩ I 11 oo co I — OI CD O O — I l UΩl COO C CDO O O CN sf 1
S CO CM S Ir CM sf 11
CO O O sf OO O O Ul CN - sf — — — CD CO l I
U) - c sf CM CD S CO CM tD CO O CD S 00 CO LO CO CD CD CD CD sf CO — CD Ul CO CO S CO CO st l LΩO — C CDD — LO O CO CO 01 sf CM CM sf O CD Ol Ol O OO OO CN Ul CD Ul st sf cO CN CD — CM LΩ Ul S L 00 00 S O — Sf CN S CM CD sf oo st LO CO CO sf O Ol LO 00 LO CD — 00 CO S S CD CN CD S — SsOoO CD) LO S s to CM Ul CD CN IO CΩ CΩ S O CN LΩ CN S CN O S O CD S — CO CD CO O CO CO
— CD Ul CO st CO O sf st st st CO sf CO CO O 00 O CD 00 CM CM 00 sf CD CD CN CO st O Ss — oO —- O CO st CD S CM — — — CO UI CO O O O CD — st CN CO S Ul st CO CO Ul CD U CO C CM sf CD tD S CD T- — — CD S ,— CD sf CO Ul CO CM sf Ul 00 S CD Ol CD CD CD — CO CO — O CO O — o O CCDD — O co CO S Ol O O CO CD CO CD Ul — Ol Ol sf Ul S - CO CO — — CM S — CD CO Ul
— CD — CO CM CN CM CO O CM CO s LΩ CO s CD CD S CO OO CD CO — CO Ul CD CD CD — CD LO — sS CCOO CCMM CCDD CO sf CD st 01 CM — S — CO sf CO CM S sf sf S CD CD st sf st S CM CM O — 00 O C CD— CO COsf CO CD COO — CD CM sf CD Ul 00 CD CD CDCO 0000 CO CDOO CO CM CN CD OICO sf CO LO CD CN — — CO sf — CD CD — CD CD — — CO sf CD CD sf — — CN sf sf CO COCO CN CM CO CO U
C) o o o o o o o o o o o o o o o o o υ υ o o o o o o o o o o υ o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o o C) Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ Φ vo φ φ 0) 0) TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ XJ XJ XJ XJ XJ XJ TJ XJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ TJ XJ XJ XJ TJ TJ TJ TJ XJ XJ TJ TJ XJ X^ o TJ TJ TJ TJ 'D — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — — r r r r r r -^ -r T -r -r -r r -r -r r -r -r i- r i- -r i^ -r -r so CO CO CO CO
CN CM CN CM CN CD CD CD CD CD CD O) σ) σ) σ) σ) σ) σ) σ) tD CD CD σi σ) CD CD CD CD σi CD CD σ) σ) CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD C^
O O O O tN CM CN CM CM tM CM tM CN CM CM CM CN CM tM eM CM CM CM CM CM CM CM CM CM tM CM tM CM CM CM CM CN CM CM CM tM CM CM tM CM
CO CO O ro CO CO CD CO CO CD CO CO CO CO CD CD CD CD CD CD CD CO CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CD CO CD CD CD CO CD CD CD CD CD CD CD CD CD
© CO 00 CO CO ∞ m co co cβ oo co co co co oo co αj oo oo co oo co αD rø oo αo αi oo co αi co cn co ∞ m o s s s s S CN CN CM CM CM CN CN CM CM CM CM CM CN CM CM CM CN CM CM CN tM CM tM eM CN tM CM CM tM CM CM tM eM IM CM CM CM CM tM CM
CD CD CD 01 CD CN CN CM CN CN CN CM CN CN CN CN CM tM CN CM CN CM CM CM CM CN CM CM CM CM CM CN CM CM CM CN CN CN CN W
CD CD CD CD CD O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O O CD CD CD CD CD S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S S
Table 4
71 0112 1.5.dec 730632H1 414 637 71 01121 1.δ.dec g2786911 1312 1665
71 0112 1.5.dec 790391 R1 429 1010 71 01121 1.δ.dec 2δ004δ2F6 1 490
71 0112 1.5.dec 790391 H1 429 671 71 01121 1.5.dec 2δ004δ2H1 1 267
71 0112 1.5.dec 4618427H1 442 687 71 01121 1.δ.dec 2604526H1 50 289
71 0112 1.5.dec 2936816H1 445 695 71 01121 1.δ.dec 2486204H1 130 370
71 0112 1.5.dec 2722677F6 470 965 71 01121 1.δ.dec 46906δδH1 205 450
71 0112 1.5.dec 1788261 H1 563 812 71 01121 1.δ.dec 1686319H1 205 396
71 0112 1.δ.dec δ116027H1 667 827 71 01121 1.δ.dec 3211206H1 205 327
71 0112 1.5.dec 5116074H1 667 821 71 01121 1.δ.dec 3411833H1 205 437
71 0112 1.5.dec 3986358H1 596 792 71 01121 1.δ.dec 3271966H1 205 464
71 0112 1.δ.dec 4203841 H1 612 909 71 01121 1.δ.dec g2161312 1211 1669
71 0112 1.δ.dec 63148δ9H1 675 816 71 01121 1.δ.dec g1368211 1227 1675
71 0112 1.δ.dec 374028δH1 1313 1599 71 01121 1.δ.dec g1801782 1241 1670
71 0112 1.δ.dec 5121062T6 1314 1660 71 01121 1.δ.dec g4163647 1241 1669
71 0112 1.δ.dec g46643δ7 1318 1671 71 01121 1.δ.dec g1317339 1256 1669
71 0112 1.δ.dec g1940230 1324 1650 71 01121 1.δ.dec 2395410T6 1256 1632
71 0112 11.δ.dec g2184646 1324 1650 71 01121 1.δ.dec 1811608T6 1262 1629
71 0112 11.δ.dec g366δδ09 1326 1676 71 01121 1.δ.dec 3488381T6 1262 1627
71 0112 H.δ.dec 2722677T6 1338 1627 71 01121 1.δ.dec g5674690 1262 1671
71 0112 11.δ.dec g2952637 1338 1673 71 01121 1.δ.dec 536468H1 1269 1522
71 0112 11.δ.dec 2474881 H1 173 407 71 01121 1.δ.dec 730632F1 1185 1671
71 0112 11.δ.dec 3580753H1 177 489 71 01121 1.δ.dec g5637214 1207 1671
71 0112 11.δ.dec 3581452H1 177 467 71 01121 1.δ.dec 1639808H1 1208 1418
71 0112 11.δ.dec 486671 H1 187 473 71 01121 1.5.dec g5669631 1208 1668
71 0112 11.δ.dec 2644119H1 194 447 71 01121 1.5.dec 2750833H1 1052 1190
71 0112 11.δ.dec 6370602H1 204 312 71 01121 1.δ.dec 2760833R6 1052 1313
71 0112 11.δ.dec 6898419H1 204 489 71 01121 1.5.dec 6176276H1 1064 1344
71 0112 11.δ.dec 3438381 H1 204 45δ 71 01121 1.5.dec 628273H1 1064 1288
71 0112 11.δ.dec 6121062F6 205 670 71 01121 1.δ.dec 2526968H1 1067 1304
71 0112 11.δ.dec 3784294H1 205 502 71 01121 1.δ.dec 59δ4055H1 1090 1214
71 0112 11.δ.dec 2107639H1 207 458 71 01121 1.δ.dec 1534862T6 1102 1626
71 0112 11.δ.dec 3732863H1 205 469 71 01121 1.δ.dec 2364626H1 1110 1325
71 0112 11.δ.dec 1721414H1 207 420 71 01121 1.δ.dec 6603769H1 1125 1673
71 0112 11.δ.dec g1367922 205 599 71 01121 1.δ.dec 6603659H1 1126 1583
71 0112 11.δ.dec 2211136H1 208 407 71 01121 1.δ.dec 6206520T6 1140 1643
71 0112 11.δ.dec 2722677H1 470 713 71 01121 1.δ.dec 780467T6 1157 1624
71 0112 1.δ.dec 1811608F6 478 916 71 01121 1.δ.dec 4837615H1 1160 1458
71 0112 1.δ.dec 1811608H1 478 721 71 01121 1.δ.dec 638428H1 1164 1311
71 0112 1.δ.dec 1809182H1 604 669 71 01121 1.δ.dec 6393337H1 813 1108
71 0112 1.δ.dec 1979630H1 610 756 71 01121 1.δ.dec 3438482H1 851 1098
71 0112 1.δ.dec 6δδ7δ38H1 516 1108 71 01121 1.δ.dec 4185657H1 871 1086
71 0112 1.δ.dec 6119348H1 522 817 71 01121 1.δ.dec 6428246H1 891 1480
71 0112 1.δ.dec 6313378H1 241 379 71 01121 1.δ.dec g2188554 894 1044
71 0112 1.δ.dec 3342386H1 243 497 71 01121 1.δ.dec 1212647H1 920 1152
71 0112 1.δ.dec 4896741 H1 252 543 71 01121 1.δ.dec 2772091 F6 929 1256
71 0112 1.δ.dec 1614966H1 229 402 71 01121 1.δ.dec 2772091 H1 929 1176
71 0112 1.δ.dec 3982780H1 234 509 71 01121 1.δ.dec 3179605H1 945 1247
71 0112 1.δ.dec 4731416H1 239 612 71 01121 1.δ.dec 5816063H1 946 1181
71 0112 1.δ.dec g2669808 1475 1671 71 01121 1.δ.dec 4691156H1 954 1182
71 0112 1.δ.dec g430731δ 1488 1669 71 01121 1.δ.dec 3411470H1 956 1210
71 0112 1.δ.dec 2662177H1 205 462 71 01121 1.δ.dec 4278241 H1 970 1200
71 0112 1.δ.dec 1866272H1 206 449 71 01121 1.δ.dec 2920077H1 1021 1284
71 0112 1.δ.dec 1866272F6 206 602 71 01121 1.5.dec 5119720H1 1027 1297
71 0112 1.δ.dec 3063770H1 220 511 71 01121 1.δ.dec 4939387H1 1037 1282
71 0112 1.δ.dec 3067168H1 220 425 71 01121 1.δ.dec 2750833T6 1041 1626
71 0112 1.δ.dec g1316022 224 666 71 01121 1.δ.dec 2500452T6 1045 1623
71 0112 1.δ.dec 2396410F6 224 572 71 01121 1.δ.dec 1904368H1 1047 1338
71 0112 1.δ.dec 239641 OH1 224 463 71 01121 1.δ.dec 1904368F6 1047 1344
71 0112 1.δ.dec 6374094H1 224 437 71 01121 1.5.dec g5640370 1392 1672
71 0112 1.5.dec g1802011 227 661 71 01121 1.δ.dec 1865272T6 1398 1839
71 0112 1.5.dec g831726 229 623 71 01121 1.δ.dec 2772091T6 1403 1625
71 0112 1.δ.dec gδ6369δ2 1271 1671 71 01121 1.δ.dec 2158516H1 1427 1547
71 0112 1.δ.dec gδδ307δδ 1282 1673 71 01121 1.δ.dec g2321614 1427 1669
71 0112 11.δ.dec gδ6δ9613 1283 1668 71 01121 1.δ.dec g4565430 1434 1669
71 0112 11.δ.dec 2589181 H1 1304 1531 71 01121 1.5.dec 790391 F1 1443 1669
71 0112 11.δ.dec g3214484 1309 1674 71 01121 1.5.dec g2021003 1472 1659
71 0112 11.δ.dec g2354244 1310 1668 71 01121 1.5.dec 4129420H1 209 384
sf xi
CO sf CD CD CM CD S CM CD IΩ CM O CN S
CM S st CO sf S st CO CN CD OO CD CD - st cO — — LΩ CM 00 CO U1 CM CD S 00 — CD O sf CN CO O O S sf CD CVl CO O CD O O O O O O CO S O O CD lΩ CD CD CD CD LΩ sf sf sf lΩ lΩ st sf CO Ul sf CD CD - — — — — — CD OO — — — — — — — —
— — CO CD CO LO CO CM CO CM CO CO sf cO CD CD Ol CD st CM CD O LO CD CM CO CN CD - S S S UO sf S — — — — — — — — — — — S CO LΩ IΩ CO CΩ CO OO OO CD O — CO CO CO CO CO CO
CM CM CM CM CM CM CM CM CM CM CM CO S S S S S S S S S OO OO — — — — — —
I σi x o r Ir I — — CD s-
O l I H O O l o s s CM CD Ul CO CN O LΩ on CD CN st CO CD CO CD sf lO CO S — S CO CN
CO CM — S 00 f CM 01 CO CO CD CN sf s O CD O CD U U sf cO S CO CM sf O st cO
CD LΩ CO CO LΩ 00 CD CO O Ul CN CO CO CN CO st st sf CD CO CO LO O CD sf st
— CD — Ul CO CN 00 CD CD CO CM CO sf CD 01 U) CD 00 CO LΩ CD st S st — sf OO
CO CD CO CD CN CD CN 00 CD 00 CO ro s cn CO LO Ul CD IO CD O O O OO LO CO CD —
01 CN S O LΩ —- CO CO CN S Sf CM LO CO st 00 st CD CO LO CO — CD CN CM OO CO CM sf LΩ S CM CO sf CM sf IO CO CM CM Ul sf CO sf CO CM st st CD— LO — CD CO COCD CD
O O O O O O O O O O O O O O O O O O O O O O O O O O O O O φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ φ vo o O O O O O T) TI O T| O TJ TI O Ti TI TJ TJ TJ O O O OO T] Tl O O Tj Ti so ib Lb ib ui ui ib Lb ui ib ui ib ibib ib ui ib ui ibui ib ui ui ui ui ui ui ui ui ui
CM CM CN CM CM CM CM CM CN CN CN CN CM CM CM CM CM CM CN CN CM CM CM CN CM CM CM CM CM
© o O O O O O O O O O O O O O O O O O O O O O O O O O O O O O
(SSSSSSSSSSSSSSSSSSSSSSSSSSSSS
TABLE 5
ID NO: Template ID Tissue Distribution
1 405310.1. oct Cardiovascular System - 16%, Germ Cells - 12%
2 48073 l .ό.oct Liver - 25%, Connective Tissue - 17%, Male Genitalia - 15%
3 334751.2.dec Germ Cells - 33%, Liver - 24%, Endocrine System - 13%
4 237330.8.dec Respiratory System - 60%, Male Genitalia - 40%
5 053778.1 1. dec Embryonic Structures - 67%, Nervous System - 17%, Digestive System -
17%
6 360645. lO.dec Embryonic Structures - 67%, Nervous System - 1 1 %
8 997089.7.dec Sense Organs - 1 1 %
9 237152.1. dec Unclassified/Mixed - 41%, Germ Cells - 19%, Embryonic Structures -
13%, Respiratory System - 13%
10 232851.7.dec Hemic and Immune System - 100%
1 1 083804.1. dec Germ Cells - 70%, Connective Tissue - 16%
12 27272 l .ό.oct Digestive System - 10%
13 461603.4.oct Germ Cells - 17%, Sense Organs - 14%, Musculoskeletal System - 10%
14 332465.2.dec Connective Tissue - 100%
15 445175.3.dec Germ Cells - 58%, Embryonic Structures - 20%, Male Genitalia - 17%
16 980541.1. dec Nervous System - 50%, Embryonic Structures - 43%
17 237996.1. dec Cardiovascular System - 21%, Exocrine Glands - 21%, Respiratory
System - 16%, Digestive System - 16%, Hemic and Immune System -
16%
19 242082. lO.dec Female Genitalia - 75%, Hemic and Immune System - 25%
20 019239.1. dec Exocrine Glands - 26%, Endocrine System - 17%, Hemic and Immune
System - 17%
21 899943.1. dec Sense Organs - 14%, Musculoskeletal System - 12%
22 443551.1. dec Unclassified/Mixed - 50%, Nervous System - 27%, Female Genitalia -
14%
23 897957.1. dec Pancreas - 44%, Liver - 27%, Male Genitalia - 10%
24 90091 1.1. dec Germ Cells - 69%, Nervous System - 12%
25 999296.1. dec Exocrine Glands - 22%, Endocrine System - 16%, Female Genitalia -
16%
26 442286.1. dec Respiratory System - 50%, Exocrine Glands - 33%, Hemic and Immune
System - 17%
27 901978.1. dec Musculoskeletal System - 25%, Unclassified/Mixed - 18%, Hemic and
Immune System - 13%
28 479346.1. dec Unclassified/Mixed - 27%, Embryonic Structures - 20%, Skin - 14%
29 481750.1. dec Male Genitalia - 29%, Urinary Tract - 13%, Skin - 1 1%
30 900917.2.dec Respiratory System - 100%
31 999415.1. dec Connective Tissue - 32%, Cardiovascular System - 32%, Exocrine
Glands - 16%
32 900680.2.dec Embryonic Structures - 34%, Liver - 19%, Unclassified/Mixed - 16%
33 902791.3.dec Urinary Tract - 72%, Nervous System - 17%, Hemic and Immune System
1 1%
34 053826.1. dec Germ Cells - 69%, Unclassified/Mixed - 22%
35 204932.4.dec NO DATA
36 400607.19.dec Musculoskeletal System - 50%, Cardiovascular System - 29%, Nervous
System - 21%
37 444248.7.dec Exocrine Glands - 57%, Digestive System - 43%
38 346599.9.dec Liver - 30%, Pancreas - 26%, Respiratory System - 21%
40 41 1396.24.dec Male Genitalia - 100%
41 302819.4.dec Exocrine Glands - 100% TABLE 5
SEQ ID NO: Template ID Tissue Distribution
42 238734.2.dec Skin - 94%
43 399525.3.dec Unclassified/Mixed - 34%, Connective Tissue - 25%, Exocrine Glands -
13%
44 222795.6.dec widely distributed
45 410628.5.dec Sense Organs - 32%, Nervous System - 12%, Urinary Tract - 1 1%,
Connective Tissue - 1 1 %
46 053649.6.dec Skin - 89%, Male Genitalia - 1 1%
47 221914.2.dec Connective Tissue - 18%, Skin - 18%, Exocrine Glands - 13%
49 401482.2.oct Skin - 12%, Connective Tissue - 10%
50 274551.1. oct Nervous System - 60%, Hemic and Immune System - 40%
51 411408.20.dec Connective Tissue - 12%, Cardiovascular System - 11%
52 035973.1. dec Embryonic Structures - 67%, Nervous System - 17%, Digestive System -
17%
53 456536.1. dec widely distributed
54 387807.4.oct Sense Organs - 33%, Skin - 1 1%, Male Genitalia - 10%
55 406790.3.dec Unclassified/Mixed - 32%, Urinary Tract - 25%, Musculoskeletal System
10%
56 412420.63.dec Sense Organs - 40%
57 196623.3.dec widely distributed
58 42791 ό.δ.dec Nervous System - 100%
59 264633.8.dec Embryonic Structures - 86%, Hemic and Immune System - 14%
61 902943.1. dec Unclassified/Mixed - 38%, Female Genitalia - 19%, Respiratory System
10%
64 197445.1. oct Unclassified/Mixed - 28%
65 348775.1. oct Skin - 28%, Connective Tissue - 15%, Nervous System - 15%
66 336239.5.dec Hemic and Immune System - 100%
67 215660.4.dec Nervous System - 100%
68 391940.2.dec Hemic and Immune System - 58%, Male Genitalia - 26%, Respiratory
System - 16%
69 978302.3.dec Sense Organs - 16%, Germ Cells - 1 1%, Embryonic Structures - 1 1%
70 228629.1 1. dec Digestive System - 60%, Hemic and Immune System - 40%
71 01 121 1.5.dec Musculoskeletal System - 32%, Endocrine System - 27%, Nervous
System - 18%
Table 6
Program Description Reference Parameter Threshold ABI FACTURA A program that removes vector sequences and masks PE Biosystems, Foster City, CA. ambiguous bases in nucleic acid sequences.
ABI PARACEL FDF A Fast Data Finder useful in comparing and annotating PE Biosystems, Foster City, CA; Mismatch <50% amino acid or nucleic acid sequences.
ABI AutoAssembler A program that assembles nucleic acid sequences. PE Biosystems, Foster City, CA.
BLAST A Basic Local Alignment Search Tool useful in sequence Altschul, S.F. et al. (1990) J. Mol. Biol. ESTs: Probability value= 1.0E-8 or less; Full similarity search for amino acid and nucleic acid 215:403-410; Altschul, S.F. et al. (1997) Length sequences: Probability value= 1.0E- sequences. BLAST includes five functions: blastp, Nucleic Acids Res. 25:3389-3402. 10 or less blastn, blastx, tblastn, and tblastx.
Is) FASTA A Pearson and Lipman algorithm that searches for Pearson, W.R. and D.J. Lipman (1988) Proc. ESTs: fasta E value=1.06E-6; Assembled
-t_ similarity between a query sequence and a group of Natl. Acad Sci. USA 85:2444-2448; ESTs: fasta Identity^ 95% or greater and sequences of the same type. FASTA comprises as least Pearson, W.R. (1990) Methods Enzymol. Match length=200 bases or greater; fastx E five functions: fasta, tfasta, fastx, tfastx, and ssearch. 183:63-98; and Smith, T.F. and M.S. value=1.0E-8 or less; Full Length sequences: Waterman (1981) Adv. Appl. Math. 2:482- fastx score=100 or greater 489.
BLIMPS A BLocks IMProved Searcher that matches a sequence Henikoff, S. and J.G. Henikoff (1991) Score=1000 or greater; Ratio of against those in BLOCKS, PRINTS, DOMO, PRODOM, Nucleic Acids Res. 19:6565-6572; Henikoff, Score/Strength = 0.75 or larger; and, if and PFAM databases to search for gene families, J.G. and S. Henikoff (1996) Methods applicable, Probability value= 1.0E-3 or less sequence homology, and structural fingeφrint regions. Enzymol. 266:88-105; and Attwood, T.K. et al. 1997 J. Chem. Inf. Com ut. Sci. 37:417-
HMMER An algorithm for searching a query sequence against Krogh, A. et al. (1994) J. Mol. Biol. Score= 10-50 bits for PFAM hits, depending hidden Markov model (HMM)-based databases of 235:1501-1531; Sonnhammer, E.L.L. et al. on individual protein families protein family consensus sequences, such as PFAM. (1988) Nucleic Acids Res. 26:320-322.
Table 6
Program Description Reference Parameter Threshold ProfileScan An algorithm that searches for structural and sequence Gribskov, M. et al. (1988) CABIOS 4:61- Normalized quality score≥GCG-specified motifs in protein sequences that match sequence patterns 66; Gribskov, M. et al. (1989) Methods "HIGH" value for that particular Prosite defined in Prosite. Enzymol. 183:146-159; Bairoch, A. et al. motif. Generally, score= 1.4-2.1.
(1997) Nucleic Acids Res. 25:217-221.
Phred A base-calling algorithm that examines automated Ewing, B. et al. (1998) Genome Res. 8:175- sequencer traces with high sensitivity and probability. 185; Ewing, B. and P. Green (1998) Genome Res. 8:186-194.
Phrap A Phils Revised Assembly Program including SWAT Smith, T.F. and M.S. Waterman (1981) Score= 120 or greater; Match length= 56 or and CrossMatch, programs based on efficient Adv. Appl. Math. 2:482-489; Smith, T.F. greater implementation of the Smith-Waterman algorithm, and M.S. Waterman (1981) J. Mol. Biol. useful in searching sequence homology and assembling 147:195-197; and Green, P., University of DNA sequences. Washington, Seattle, WA. -t_
4-
Consed A graphical tool for viewing and editing Phrap Gordon, D. et al. (1998) Genome Res. 8:195- assemblies. 202.
SPScan A weight matrix analysis program that scans protein Nielson, H. et al. (1997) Protein Engineering Score=3.5 or greater sequences for the presence of secretory signal peptides. 10: 1-6; Claverie, J.M. and S. Audic (1997) CABIOS 12:431-439.
Motifs A program that searches amino acid sequences for Bairoch, A. et al. (1997) Nucleic Acids Res. patterns that matched those defined in Prosite. 25:217-221; Wisconsin Package Program Manual, version 9, page M51-59, Genetics Computer Group, Madison, WI.

Claims

CLAIMS What is claimed is:
1. An isolated polynucleotide comprising a polynucleotide sequence selected from the group consisting of: a) a polynucleotide sequence selected from the group consisting of SEQ ID NO: 1-71, b) a naturally occurring polynucleotide sequence having at least 90% sequence identity to a polynucleotide sequence selected from the group consisting of SEQ ID NO: 1-71, c) a polynucleotide sequence complementary to a), d) a polynucleotide sequence complementary to b), and e) an RNA equivalent of a) through d).
2. An isolated polynucleotide of claim 1, comprising a polynucleotide sequence selected from the group consisting of SEQ ID NO : 1 -71.
3. An isolated polynucleotide comprising at least 60 contiguous nucleotides of a polynucleotide of claim 1.
4. A composition for the detection of expression of diagnostic and therapeutic polynucleotides comprising at least one of the polynucleotides of claim 1 and a detectable label.
5. A method for detecting a target polynucleotide in a sample, said target polynucleotide having a sequence of a polynucleotide of claim 1 , the method comprising: a) amplifying said target polynucleotide or fragment thereof using polymerase chain reaction ampUfication, and b) detecting the presence or absence of said amplified target polynucleotide or fragment thereof, and, optionally, if present, the amount thereof.
6. A method for detecting a target polynucleotide in a sample, said target polynucleotide comprising a sequence of a polynucleotide of claim 1 , the method comprising: a) hybridizing the sample with a probe comprising at least 20 contiguous nucleotides comprising a sequence complementary to said target polynucleotide in the sample, and which probe specifically hybridizes to said target polynucleotide, under conditions whereby a hybridization complex is formed between said probe and said target polynucleotide or fragments thereof, and b) detecting the presence or absence of said hybridization complex, and, optionally, if present, the amount thereof.
7. A method of claim 5, wherein the probe comprises at least 30 contiguous nucleotides.
8. A method of claim 5, wherein the probe comprises at least 60 contiguous nucleotides.
9. A recombinant polynucleotide comprising a promoter sequence operably linked to a polynucleotide of claim 1.
10. A cell transformed with a recombinant polynucleotide of claim 9.
11. A transgenic organism comprising a recombinant polynucleotide of claim 9.
12. A method for producing a diagnostic and therapeutic polypeptide, the method comprising: a) culturing a cell under conditions suitable for expression of the diagnostic and therapeutic polypeptide, wherein said cell is transformed with a recombinant polynucleotide of claim 9, and" b) recovering the diagnostic and therapeutic polypeptide so expressed.
13. A purified diagnostic and therapeutic polypeptide encoded by at least one of the polynucleotides of claim 2.
14. An isolated antibody which specifically binds to a diagnostic and therapeutic polypeptide of claim 13.
15. A method of identifying a test compound which specifically binds to the diagnostic and therapeutic polypeptide of claim 13, the method comprising the steps of: a) providing a test compound; b) combining the diagnostic and therapeutic polypeptide with the test compound for a sufficient time and under suitable conditions for binding; and c) detecting binding of the diagnostic and therapeutic polypeptide to the test compound, thereby identifying the test compound which specifically binds the diagnostic and therapeutic polypeptide.
16. A microarray wherein at least one element of the microarray is a polynucleotide of claim 3.
17. A method for generating a transcript image of a sample which contains polynucleotides, the method comprising the step's of: a) labeling the polynucleotides of the sample, b) contacting the elements of the microarray of claim 16 with the labeled polynucleotides of the sample under conditions suitable for the formation of a hybridization complex, and c) quantifying the expression of the polynucleotides in the sample.
18. A method for screening a compound for effectiveness in altering expression of a target polynucleotide, wherein said target polynucleotide comprises a polynucleotide sequence of claim 1, the method comprising: a) exposing a sample comprising the target polynucleotide to a compound, under conditions suitable for the expression of the target polynucleotide, b) detecting altered expression of the target polynucleotide, and c) comparing the expression of the target polynucleotide in the presence of varying amounts of the compound and in the absence of the compound.
19. A method for assessing toxicity of a test compound, said method comprising: a) treating a biological sample containing nucleic acids with the test compound; b) hybridizing the nucleic acids of the treated biological sample with a probe comprising at least 20 contiguous nucleotides of a polynucleotide of claim 1 under conditions whereby a specific hybridization complex is formed between said probe and a target polynucleotide in the biological sample, said target polynucleotide comprising a polynucleotide sequence of a polynucleotide of claim 1 or fragment thereof; c) quantifying the amount of hybridization complex; and d) comparing the amount of hybridization complex in the treated biological sample with the amount of hybridization complex in an untreated biological sample, wherein a difference in the amount of hybridization complex in the treated biological sample is indicative of toxicity of the test compound.
EP00963614A 1999-09-23 2000-09-19 Molecules for diagnostics and therapeutics Withdrawn EP1224275A2 (en)

Applications Claiming Priority (49)

Application Number Priority Date Filing Date Title
US15576099P 1999-09-23 1999-09-23
US155760P 1999-09-23
US15629499P 1999-09-24 1999-09-24
US15593999P 1999-09-24 1999-09-24
US156294P 1999-09-24
US155939P 1999-09-24
US15656599P 1999-09-28 1999-09-28
US15662599P 1999-09-28 1999-09-28
US15662499P 1999-09-28 1999-09-28
US156624P 1999-09-28
US156565P 1999-09-28
US156625P 1999-09-28
US16752299P 1999-11-24 1999-11-24
US16751799P 1999-11-24 1999-11-24
US16741099P 1999-11-24 1999-11-24
US16745399A 1999-11-24 1999-11-24
US16752199P 1999-11-24 1999-11-24
US16752099P 1999-11-24 1999-11-24
US16754299P 1999-11-24 1999-11-24
US167520P 1999-11-24
US167410P 1999-11-24
US167453P 1999-11-24
US167542P 1999-11-24
US167521P 1999-11-24
US167517P 1999-11-24
US167522P 1999-11-24
US16794599P 1999-11-29 1999-11-29
US16794399P 1999-11-29 1999-11-29
US167945P 1999-11-29
US167943P 1999-11-29
US16819799P 1999-11-30 1999-11-30
US16842999P 1999-11-30 1999-11-30
US16826599P 1999-11-30 1999-11-30
US16843299P 1999-11-30 1999-11-30
US168429P 1999-11-30
US168432P 1999-11-30
US168265P 1999-11-30
US168197P 1999-11-30
US16846899P 1999-12-01 1999-12-01
US16859999P 1999-12-01 1999-12-01
US168468P 1999-12-01
US168599P 1999-12-01
US16885799P 1999-12-02 1999-12-02
US16861399P 1999-12-02 1999-12-02
US16861199P 1999-12-02 1999-12-02
US168611P 1999-12-02
US168857P 1999-12-02
US168613P 1999-12-02
PCT/US2000/025643 WO2001021836A2 (en) 1999-09-23 2000-09-19 Molecules for diagnostics and therapeutics

Publications (1)

Publication Number Publication Date
EP1224275A2 true EP1224275A2 (en) 2002-07-24

Family

ID=27586738

Family Applications (1)

Application Number Title Priority Date Filing Date
EP00963614A Withdrawn EP1224275A2 (en) 1999-09-23 2000-09-19 Molecules for diagnostics and therapeutics

Country Status (1)

Country Link
EP (1) EP1224275A2 (en)

Non-Patent Citations (1)

* Cited by examiner, † Cited by third party
Title
See references of WO0121836A2 *

Similar Documents

Publication Publication Date Title
US20040115629A1 (en) Molecules for diagnostics and therapeutics
WO2004023973A2 (en) Molecules for diagnostics and therapeutics
CA2442705A1 (en) Molecules for diagnostics and therapeutics
US20040014087A1 (en) Molecules for diagnostics and therapeutics
JP2004528003A (en) Extracellular matrix and cell adhesion molecules
WO2000073509A2 (en) Molecules for diagnostics and therapeutics
US20040048253A1 (en) Molecules for diagnostics and therapeutics
JP2004516812A (en) Receptor
JP2003529325A (en) Human transport protein
JP2004500114A (en) Transcription factor
WO2003062376A2 (en) Molecules for diagnostics and therapeutics
WO2001062927A2 (en) Polypeptides and corresponding polynucleotides for diagnostics and therapeutics
EP1265998A2 (en) Polypeptides and corresponding polynucleotides for diagnostics and therapeutics
JP2003532419A (en) Cytoskeletal binding protein
EP1364015A2 (en) Molecules for diagnostics and therapeutics
WO2001021836A2 (en) Molecules for diagnostics and therapeutics
JP2003517290A (en) Human transcription regulatory protein
WO2003062385A2 (en) Secretory molecules
WO2002079473A2 (en) Molecules for diagnostics and therapeutics
JP2004511208A (en) RNA metabolism protein
JP2005508636A (en) Nucleic acid binding protein
JP2005500008A (en) Receptors and membrane-bound proteins
JP2004528002A (en) Secretory and transport molecules
US20040171012A1 (en) Nucleic acid-associated proteins
JP2004509610A (en) Nuclear hormone receptor

Legal Events

Date Code Title Description
PUAI Public reference made under article 153(3) epc to a published international application that has entered the european phase

Free format text: ORIGINAL CODE: 0009012

17P Request for examination filed

Effective date: 20020419

AK Designated contracting states

Kind code of ref document: A2

Designated state(s): AT BE CH CY DE DK ES FI FR GB GR IE IT LI LU MC NL PT SE

AX Request for extension of the european patent

Free format text: AL;LT;LV;MK;RO;SI

RIN1 Information on inventor provided before grant (corrected)

Inventor name: COHEN, HOWARD J

Inventor name: BRATCHER; SHAWN R

Inventor name: HODGSON,DAVID, M.

Inventor name: BANVILLE, STEVEN C

Inventor name: DUFOUR, GERARD F

Inventor name: SPIRO, PETER A.

Inventor name: RUSSO, FRANK D

Inventor name: LINCOLN, STEPHEN E

RIN1 Information on inventor provided before grant (corrected)

Inventor name: COHEN, HOWARD J

Inventor name: LINCOLN, STEPHEN E

Inventor name: DUFOUR, GERARD F

Inventor name: HODGSON,DAVID, M.

Inventor name: BRATCHER; SHAWN R

Inventor name: BANVILLE, STEVEN C

Inventor name: SPIRO, PETER A.

Inventor name: RUSSO, FRANK D

STAA Information on the status of an ep patent application or granted ep patent

Free format text: STATUS: THE APPLICATION HAS BEEN WITHDRAWN

18W Application withdrawn

Effective date: 20050125

RIN1 Information on inventor provided before grant (corrected)

Inventor name: DUFOUR, GERARD F

Inventor name: LINCOLN, STEPHEN E

Inventor name: CHALUP, MICHAEL S

Inventor name: STOCKDREHER, THERESA K

Inventor name: ROSEN, BRUCE H

Inventor name: GREENAWALT, LILA B

Inventor name: CHEN, WENSHENG

Inventor name: AMSHEY, STEFAN

Inventor name: YU, JIMMY Y

Inventor name: SPIRO, PETER A.

Inventor name: HODGSON,DAVID, M.

Inventor name: ROSEBERRY, ANN M

Inventor name: JACKSON, JENNIFER L

Inventor name: RUSSO, FRANK D

Inventor name: LIU, TOMMY,F

Inventor name: FONG, WILLY T

Inventor name: YAP, PIERRE E

Inventor name: BRATCHER; SHAWN R

Inventor name: BANVILLE, STEVEN C

Inventor name: JONES, ANISSA LEE

Inventor name: WRIGHT, RACHEL J

Inventor name: COHEN, HOWARD J

Inventor name: PANZER, SCOTT R

Inventor name: SHAH, PURVI