EP4680631A1 - Combotope antibody libraries - Google Patents

Combotope antibody libraries

Info

Publication number
EP4680631A1
EP4680631A1 EP24707819.9A EP24707819A EP4680631A1 EP 4680631 A1 EP4680631 A1 EP 4680631A1 EP 24707819 A EP24707819 A EP 24707819A EP 4680631 A1 EP4680631 A1 EP 4680631A1
Authority
EP
European Patent Office
Prior art keywords
antibody
domain
amino acid
seq
library
Prior art date
Legal status (The legal status is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the status listed.)
Pending
Application number
EP24707819.9A
Other languages
German (de)
French (fr)
Inventor
Ola Blixt
Ramón HURTADO GUERRERO
Spyridon GATOS
Current Assignee (The listed assignees may be inaccurate. Google has not performed a legal analysis and makes no representation or warranty as to the accuracy of the list.)
Danmarks Tekniske Universitet
Universidad de Zaragoza
Fundacion Agencia Aragonesa para la Investigacion y el Desarrollo ARAID
Original Assignee
Danmarks Tekniske Universitet
Universidad de Zaragoza
Fundacion Agencia Aragonesa para la Investigacion y el Desarrollo ARAID
Priority date (The priority date is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the date listed.)
Filing date
Publication date
Priority claimed from PCT/EP2023/056922 external-priority patent/WO2024193794A1/en
Application filed by Danmarks Tekniske Universitet, Universidad de Zaragoza, Fundacion Agencia Aragonesa para la Investigacion y el Desarrollo ARAID filed Critical Danmarks Tekniske Universitet
Publication of EP4680631A1 publication Critical patent/EP4680631A1/en
Pending legal-status Critical Current

Links

Classifications

    • CCHEMISTRY; METALLURGY
    • C07ORGANIC CHEMISTRY
    • C07KPEPTIDES
    • C07K16/00Immunoglobulins [IG], e.g. monoclonal or polyclonal antibodies
    • C07K16/005Immunoglobulins [IG], e.g. monoclonal or polyclonal antibodies constructed by phage libraries
    • CCHEMISTRY; METALLURGY
    • C07ORGANIC CHEMISTRY
    • C07KPEPTIDES
    • C07K16/00Immunoglobulins [IG], e.g. monoclonal or polyclonal antibodies
    • C07K16/18Immunoglobulins [IG], e.g. monoclonal or polyclonal antibodies against material from animals or humans
    • C07K16/28Immunoglobulins [IG], e.g. monoclonal or polyclonal antibodies against material from animals or humans against receptors, cell surface antigens or cell surface determinants
    • C07K16/2896Immunoglobulins [IG], e.g. monoclonal or polyclonal antibodies against material from animals or humans against receptors, cell surface antigens or cell surface determinants against molecules with a "CD"-designation, not provided for elsewhere
    • CCHEMISTRY; METALLURGY
    • C07ORGANIC CHEMISTRY
    • C07KPEPTIDES
    • C07K16/00Immunoglobulins [IG], e.g. monoclonal or polyclonal antibodies
    • C07K16/18Immunoglobulins [IG], e.g. monoclonal or polyclonal antibodies against material from animals or humans
    • C07K16/28Immunoglobulins [IG], e.g. monoclonal or polyclonal antibodies against material from animals or humans against receptors, cell surface antigens or cell surface determinants
    • C07K16/30Immunoglobulins [IG], e.g. monoclonal or polyclonal antibodies against material from animals or humans against receptors, cell surface antigens or cell surface determinants from tumour cells
    • C07K16/3076Immunoglobulins [IG], e.g. monoclonal or polyclonal antibodies against material from animals or humans against receptors, cell surface antigens or cell surface determinants from tumour cells against structure-related tumour-associated moieties
    • C07K16/3092Immunoglobulins [IG], e.g. monoclonal or polyclonal antibodies against material from animals or humans against receptors, cell surface antigens or cell surface determinants from tumour cells against structure-related tumour-associated moieties against tumour-associated mucins
    • CCHEMISTRY; METALLURGY
    • C07ORGANIC CHEMISTRY
    • C07KPEPTIDES
    • C07K2317/00Immunoglobulins specific features
    • C07K2317/20Immunoglobulins specific features characterized by taxonomic origin
    • C07K2317/24Immunoglobulins specific features characterized by taxonomic origin containing regions, domains or residues from different species, e.g. chimeric, humanized or veneered
    • CCHEMISTRY; METALLURGY
    • C07ORGANIC CHEMISTRY
    • C07KPEPTIDES
    • C07K2317/00Immunoglobulins specific features
    • C07K2317/30Immunoglobulins specific features characterized by aspects of specificity or valency
    • C07K2317/31Immunoglobulins specific features characterized by aspects of specificity or valency multispecific
    • CCHEMISTRY; METALLURGY
    • C07ORGANIC CHEMISTRY
    • C07KPEPTIDES
    • C07K2317/00Immunoglobulins specific features
    • C07K2317/30Immunoglobulins specific features characterized by aspects of specificity or valency
    • C07K2317/34Identification of a linear epitope shorter than 20 amino acid residues or of a conformational epitope defined by amino acid residues
    • CCHEMISTRY; METALLURGY
    • C07ORGANIC CHEMISTRY
    • C07KPEPTIDES
    • C07K2317/00Immunoglobulins specific features
    • C07K2317/50Immunoglobulins specific features characterized by immunoglobulin fragments
    • C07K2317/56Immunoglobulins specific features characterized by immunoglobulin fragments variable (Fv) region, i.e. VH and/or VL
    • CCHEMISTRY; METALLURGY
    • C07ORGANIC CHEMISTRY
    • C07KPEPTIDES
    • C07K2317/00Immunoglobulins specific features
    • C07K2317/50Immunoglobulins specific features characterized by immunoglobulin fragments
    • C07K2317/56Immunoglobulins specific features characterized by immunoglobulin fragments variable (Fv) region, i.e. VH and/or VL
    • C07K2317/565Complementarity determining region [CDR]
    • CCHEMISTRY; METALLURGY
    • C07ORGANIC CHEMISTRY
    • C07KPEPTIDES
    • C07K2317/00Immunoglobulins specific features
    • C07K2317/60Immunoglobulins specific features characterized by non-natural combinations of immunoglobulin fragments
    • C07K2317/62Immunoglobulins specific features characterized by non-natural combinations of immunoglobulin fragments comprising only variable region components
    • C07K2317/622Single chain antibody (scFv)
    • CCHEMISTRY; METALLURGY
    • C07ORGANIC CHEMISTRY
    • C07KPEPTIDES
    • C07K2317/00Immunoglobulins specific features
    • C07K2317/90Immunoglobulins specific features characterized by (pharmaco)kinetic aspects or by stability of the immunoglobulin
    • C07K2317/92Affinity (KD), association rate (Ka), dissociation rate (Kd) or EC50 value
    • CCHEMISTRY; METALLURGY
    • C12BIOCHEMISTRY; BEER; SPIRITS; WINE; VINEGAR; MICROBIOLOGY; ENZYMOLOGY; MUTATION OR GENETIC ENGINEERING
    • C12NMICROORGANISMS OR ENZYMES; COMPOSITIONS THEREOF; PROPAGATING, PRESERVING, OR MAINTAINING MICROORGANISMS; MUTATION OR GENETIC ENGINEERING; CULTURE MEDIA
    • C12N15/00Mutation or genetic engineering; DNA or RNA concerning genetic engineering, vectors, e.g. plasmids, or their isolation, preparation or purification; Use of hosts therefor
    • C12N15/09Recombinant DNA-technology
    • C12N15/10Processes for the isolation, preparation or purification of DNA or RNA
    • C12N15/1034Isolating an individual clone by screening libraries
    • C12N15/1037Screening libraries presented on the surface of microorganisms, e.g. phage display, E. coli display

Definitions

  • the invention concerns antibodies, antibody libraries and methods for identifying antibodies, which target Tn- and STn- glycosylation site of any protein site of choice, especially relevant for binding to cancer cell targets.
  • the antiboides identified by the new concept proposed herein have specificity towards both the sugar epitope as well as the peptide backbone in a glycoprotein associated with or carrying the sugar epitope on a cancer cell.
  • the invention provides specific Tn- and STn antibody libraries and methods for identifying specific antibodies, which target Tn- or STn- glycosylation sites of any glycoprotein of choice, especially relevant for binding to cancer cell targets, such as tumor cells.
  • the invention further provides antibodies identified by the new concept proposed herein, which have combined specificity towards both the carbohydrate epitope as well as the peptide epitope of the glycoprotein.
  • a dense layer of complex carbohydrate structures covers almost all eukaryotic cells. Tumor cells, contrary to their healthy counterparts, exhibit altered glycosylation patterns on the cell surface. Such altered cancer glycosylation includes increased sialylation, fucosylation, short truncated O-glycans and increased N-branching.
  • TACAs tumor-associated carbohydrate antigens
  • Tn and STn are not commonly observed in any normal human or rodent tissues, but are highly expressed on many solid tumors/carcinomas. Thus, Tn and STn represent major targets of potential immunotherapy as well as being useful in diagnostic of cancereous states.
  • MUC1 is the most well-studied mucin from the mucin family and is present in many adenocarcinomas displaying short truncated O-glycans. Under healthy conditions, MUC1 peptide core is heavily glycosylated and therefore masked by the O-glycan moieties that protect MUC1 from proteolytic cleavage enzymes. In adenocarcinomas, MUC1 proteins have shorter and less dense O-glycan side chains, resulting in exposure of the core domains of the protein on the cell surface. This altered glycosylation on MUC1 results in exposure of the epitopes MUCl-Tn and MUCl-STn to the immune system.
  • CD43 (leukosialin) is a type I transmembrane sialoglycoprotein that is abundant in hematopoietic cells, including lymphocytes, monocytes, granulocytes, natural killer cells, platelets except resting mature B cells, and erythrocytes.
  • the human CD43 protein has a mucin-type extracellular domain rich in serine and threonine residues enabling extensive O-GalNAc glycosylation with significant molecular weight heterogeneity.
  • CD43 glycoforms have been reported in several hematological and non-hematopoietic cancers, including the lung, breast, colon, cervix, and prostate, which express CD43 mostly in the early stages of tumor progression.
  • Tn/STn Known antibodies releated to Tn/STn include 5E5 (Macias-Leon et al 2020; Tarp et al 2007; Blixt et al 2010), anti-CD43 (Blixt et al 2012), 2D9 (Sorensen et al 2006; Tarp et al 2007; Blixt et al 2010), and 5F7 (US11161911B2).
  • Tn/STn Further known antibodies releating to Tn/STn include G2D11 (Persson et al 2017), 3F1 (Kjeldsen et al 1988), 83D4 (Oppezzo et al 2004), 15G9 (Mazal et al 2013), 1E3 (Li et al 2009), and MLS128 (Yuasa et al 2012); 16E12.1D9.1B11 (WO 2023/034569 Al), and 1 A5-2C9 (US 2022/057402 Al), however, these antibodies are Tn-hapten binders with no or unknown binding contribution to the peptide/protein carrier, as will be further discussed herein.
  • Monoclonal antibodies to the Tn and STn antigens are notably difficult to generate and are expensive to produce, and their specificities are often not well characterized, especially in regard to whether these antibodies simultaneously recognize the Tn/STn and the protein backbone/carrier, a requirement for superior specificity and therapeutic use.
  • the present invention provides a new antibody concept technology for rapid development of combotope antibodies (Abs) targeting Tn- and STn- glycosylation site of any glycoprotein site of choice, paving the way for a new generation of therapeutic opportunities in cancer treatment and diagnostics.
  • a novel structural and biochemical explanation, provided herein for the first time, of Tn- and STn- recognition specifically by the VH hypervariable region of the antibody, is combined with phage display screening for specific peptide recognition by the VL domains, thereby providing combotope antibodies with high speficicity and affinity to desired biological glycoprotein targets due to the combined recognition of the Tn and/or STn carbohydrate epitope and the protein backbone/carrier associated with the Tn and/or STn carbohydrate epitope.
  • Phage display is a fast method for antibody development.
  • the constructed Tn and STn template libraries define the VH part of the antibody binding the desired glycoform, and the VL diversity will determine the peptide backbone specificity.
  • MUC1 was used as a proof of concept target for both libraries (Tn and STn) to evaluate the functionality of these two libraries.
  • Single chain variant fragment (scFv)s with sequences being the same as or similar to the already known MUC1 antibodies were isolated from biopanning of the libraries.
  • CD43 was further the first target that was used to identify scFvs against Tn-CD43 peptide.
  • Tn-MUCl and Tn-CD43 scFvs were characterized for their specificity in vitro, and the scFvs demonstrated their potential uses as therapeutics on cancer cell lines. Furthermore, STn-MUCl scFvs are identidifed.
  • the constructed libraries are the first libraries to be specifically designed for glycosylated targets.
  • the constructed libraries are the first libraries to be specifically designed for specific glycosylated targets defined by the specific glycoprotein bearing the short truncated O-glycans (Tn or STn) presented on the surface of many cancer cells.
  • Tn or STn short truncated O-glycans
  • the invention provides an antibody library for in-vitro identification of a specific antibody which binds a tumor cell, wherein each antibody in said library comprises
  • a second antibody domain selected from a repertoire of second antibody domains, wherein the repertoire of second antibody domains comprises one or more second antibody domains which binds a peptide epitope of said glycoprotein of said tumor cell, wherein said specific antibody is specific for a combination of said carbohydrate epitope and said peptide epitope of said glycoprotein.
  • the first antibody domain is a VH-domain and the second antibody domain is a VL-domain.
  • the invention provides an antibody library, wherein each antibody in the library comprises
  • VL-domain selected from a repertoire of VL-domains, wherein the repertoire of VL-domains comprises one or more VL-domains which binds a peptide epitope of said glycoprotein of said tumor cell, for in-vitro identification of a specific antibody from said library which binds a tumor cell; wherein the specific antibody is specific for a combination of said carbohydrate epitope and said peptide epitope of said glycoprotein.
  • the invention provides a nucleic acid library encoding the antibody library of the first aspect of the invention.
  • the invention provides a method for identifying an antibody for targeting a tumor cell, comprising the steps of i) preparing an antibody library according to the first aspect of the invention, and ii) screening said library to identify one or more tumor targeting antibodies, preferably one or more specific tumor targeting antibodies.
  • the invention provides a method for identifying a glycopeptide target, said target comprising a Tn and/or STn epitope and a peptide epitope, such as a glycoprotein target of a cancer cell, said method comprising the steps of i) preparing an antibody library according to the first aspect of the invention, and ii) incubating the antibody library with a sample comprising the glycopeptide target, iii) analyzing one or more antibody-peptide complexes obtained from step ii) to identify the amino acid sequence of the peptide epitope of said glycopeptide target.
  • the invention provides a specific tumor cell binding antibody, comprising
  • the antibodies may be used in method of treatment and/or prevention of cancer.
  • the antibodies may be used in diagnosing cancerous states in vivo or ex vivo in cell samples.
  • the antibodies of the invention as described herein may also be subjected to one or more techniques known per se for improving one or more desired properties of the antibodies, such as (improved) affinity, (improved) potency or (reduced) immunogeneticy, and such improved antibodies form further aspects of the invention.
  • the antibodies may be subjected to techniques for affinity maturation known per se.
  • potential immunogenicity may be reduced or removed by techniques for humanization known per se and/or using techniques for identifying potential immunogenic epitopes and then removing said potential epitopes by means of one or more suitable amino acid mutations, which again can be performed in a manner known per se.
  • the amino acid of the antibody (or the sequence encoding such antibody) may for example also be subjected to techniques known per se for providing improved and/or increased expression in a desired host cell or host organism to be used for expression and/or production, which again will be clear to the skilled person, and may include known techniques for codon optimization.
  • each of the preceding techniques may require or include a limited degree of trial-and-error which will be well within the skill of the artisan.
  • a library as described herein will usually comprise at least 10 different antibodies/clones or more, such as at least 50 different antibodies/clones or more, for example at least 100 different antibodies/clones or more.
  • the upper limit for the size of a library as described herein is not critical and may be determined more by considerations such as desired diversity and practical considerations such as the size of library that can be easy and/or conveniently to generate, handle and screen.
  • a library as described herein may comprise more than 10 4 different clones, for example more than 10 5 different clones, such as more than 10 7 different clones and even up to 10 8 , 10 9 , IO 10 or more different clones.
  • a library as used herein will comprise at least 1000 (10 3 ), such as at least 10.000 (10 4 ), preferably at least 100.000 (10 5 ), such as 10 5 or more, different sequences (i.e. antibodies with different VH/VL pairs).
  • generating such library will generally comprise combining a VH sequence that has been chosen for its ability to specifically bind to a Tn or STn epitope (such as one of the VH sequences described herein or referred to herein) with a repertoire of VL sequences that provides the desired size and diversity to the library.
  • a VL repertoire may be a collection of naive (for example obtained from a naive library generated from mouse B-cells or human B-cells), pre-immune, synthetic or semisynthetic sequences.
  • such a library may be in any suitable format, including DNA, RNA or protein, and may be in the form of a library in which the VH and VL sequences are present in a suitable format/vector that allows for suitable expression or display of the antibodies.
  • This may for example be a phage library or yeast library, depending on the technique(s) that are intended to be used for screening the library.
  • Suitable screening techniques and suitable library formats for use in such screening techniques will be clear to the skilled person, and for example and without limitation include (techniques and libraries for) phage display, ribosome display, yeast display or display using suitable mammalian cell systems.
  • CDR CDR1
  • CDR2 CDR3
  • amino acid difference refers to an insertion, deletion or substitution of a single amino acid residue on a position of the first sequence, compared to the second sequence; it being understood that two amino acid sequences can contain one, two or more such amino acid differences;
  • binding when referring to the ability of an antibody to bind to an antigen, such binding is preferably specific binding, which typically means that such an antibody will bind to its antigen with a dissociation constant (KD) of 10 -5 to 10 12 moles/liter or less, and preferably IO -7 to 10 12 moles/liter or less and more preferably 10 -8 to 10 12 moles/liter (i.e. with an association constant (KA) of 10 5 to 10 12 liter/ moles or more, and preferably 10 7 to 10 12 liter/moles or more and more preferably 10 8 to 10 12 liter/moles).
  • KD dissociation constant
  • KA association constant
  • any KD value greater than 10 4 mol/liter (or any KA value lower than 10 4 M-l) liters/mol is generally considered to indicate non-specific binding.
  • an antibody of the invention will bind to the desired antigen with an affinity less than 500 nM, preferably less than 200 nM, more preferably less than 10 nM, such as less than 500 pM.
  • Specific binding of an antigen-binding protein to an antigen or antigenic determinant can be determined in any suitable manner known per se, including, for example, Scatchard analysis and/or competitive binding assays, such as radioimmunoassays (RIA), enzyme immunoassays (EIA) and sandwich competition assays, and the different variants thereof known per se in the art; as well as the other techniques mentioned herein.
  • Scatchard analysis and/or competitive binding assays such as radioimmunoassays (RIA), enzyme immunoassays (EIA) and sandwich competition assays, and the different variants thereof known per se in the art; as well as the other techniques mentioned herein.
  • VH domains and VL domains referred to herein will be part of, and in some cases in essence may have to be part of, a VH/VL pair in order to form a complete binding site. It will also be clear to the skilled person that, for this reason, it may not in all cases be practicable or possible to determine whether an individual VH domain or VL domain can bind to an antigen or epitope when it is not part of a VH/VL pair.
  • a VH domain or VL domain is referred to as being able to (specifically) bind to an epitope, antigen or protein, this includes both a situation where such domain can bind (and/or such binding can be determined or measured) when the domain is in isolated form (i.e. not part of a VH/VL pair) or where binding of such a domain takes place (and/or can be determined or measured) when such a domain is part of a suitable VH/VL pair as described herein (i.e. a VH/VL pair as is present in the antibodies described herein).
  • the invention provides a method for generating antibodies against a target that is present on the surface of a cancer cell (and/or is expressed by a cancer cell on its surface), in which said target is present on/expressed on the cancer cell in a glycosylated form (and in particular in a glycosylated form that contains, comprises or presents one or more Tn ot STn epitopes, as as described herein), which method at least comprises the following steps:
  • each antibody in the library comprises a VH domain and VL domain, in which each such VH domain is (chosen to be) capable of binding (and preferably specifically binding, as further defined herein) to a Tn epitope or a STn epitope
  • the library is preferably such that it comprises (a collection or repertoire of) at least 1000 (10 3 ), such as at least 10.000 (10 4 ), preferably at least 100.000 (IO 5 ), such as IO 5 or more, different VL domains such that, upon each of said VL domains combining with said VH domain as part of an antibody that can be obtained from said library, said library comprises/provides (a collection or repertoire of) at least 1000 (10 3 ), such as at least 10.000 (10 4 ), preferably at least 100.000 (10 5 ), such as 10 5 or more, different antibodies (i.e.
  • antibodies having different VH/VL combinations screening said library against said target, and preferably screening said library against said target in which said target is in a glycosylated form, and even more preferably screening said library against said target in which said target is in a glycosylated form that contains, comprises or presents one or more Tn or STn epitopes; identifying, generating and/or isolating one or more antibodies (or sequences encoding such antibodies) that are capable of specifically binding to said target, and that are preferably capable of specifically binding to said target in a glycosylated form, and more preferably capable of specifically binding to said target in a glycosylated form that contains, comprises or presents one or more Tn or STn epitopes; and optionally applying one or more steps for improving one or desired properties of the antibody/antibodies thus obtained, for example by means of humanization, affinity maturation, removing potential immunogenic epitopes and/or optimizing the sequence for expression or production in a desired host cell or host organism.
  • some of the libraries of the invention may comprise a single VH sequence (i.e. chosen to specifically bind to a Tn epitope or STn epitope, as further described herein) that is combined with a suitable collection or repertoire of different VL sequences, for example at least 1000 (10 3 ), such as at least 10.000 (10 4 ), preferably at least 100.000 (10 5 ), such as 10 5 or more different VL sequences.
  • VH sequences may contain two or more different VH sequences, in which each such VH sequence has been chosen to specifically bind to a Tn epitope or to an STn epitope (in which said VH sequences are again suitably combined with a suitable collection or repertoire of VL sequences, as further described herein).
  • VH sequences that can be used to generate the libraries of the invention described herein will be clear to the skilled person based on the disclosure herein.
  • suitable candidates for VH sequences include, but are not limited to, VH sequences that are present in and/or have been derived from antibodies that have been raised against glycosylated proteins that are present and/or expressed on the surface of cancer cells, in particular where such proteins are known to contain, comprise or present Tn or STn epitopes. Such VH sequences may then be tested for their ability to be used as a VH sequence in constructing the libraries described herein.
  • the VH sequences that are present in the antibodies 5E5, anti-CD43, 2D9, 5F7, G2D11, 3F1, 83D4, 15G9, 1E3, MLS128, 16E12.1D9.1B11 and/or 1A5-2C9 may be tested for their ability to serve as a VH sequence in the libraries of the invention (and, by extension, in antibodies generated from such a library); and such VH sequences (or VH sequences comprising the CDRs that are present in these antibodies) may then be used as VH sequences in constructing the libraries provided by the invention.
  • the libraries of the invention are preferably such that all or essentially all VH sequences that are present in the library (and/or that have been used in constructing the library) are capable of binding to either a Tn or STn epitope. Furthermore, the library of the invention is also preferably such that at least 90%, such as at least 95%, and more preferably all or essentially all antibodies that are obtained by means of the methods described herein (i.e.
  • VH sequences that is capable of binding to a Tn or STn epitope; and even more preferably contain a VH sequence that confers, to the antibody or antibodies obtained from the library, the ability to specifically bind to (the TN or STn part) of an epitope or antigen that comprises a Tn or STn epitope.
  • the invention also provides some VH sequences that have been found to be particularly suited for use as VH sequences in the libraries provided by the invention (and again, by extension, in antibodies that can be generated using such libraries). These are the VH sequence of SEQ ID NO: 1 (which can be used to generate libraries that have specificity for Tn epitopes/proteins comprising Tn epitopes) and the VH sequence of SEQ ID NO: 28 (which can be used to generate libraries that have specificity for STn epitopes/proteins comprising STn epitopes).
  • VH sequences and other suitable VH sequences having the same CDRs as the VH sequences of SEQ ID NO: 1 or SEQ ID NO: 28, respectively
  • libraries of the invention comprising and/or based on such VH sequences, methods for generating such librariesm and antibodies identified using, generated using and/or isolated from such libraries form further preferred aspects of the invention.
  • VH sequences will be clear to the skilled person based on the disclosure herein and include: (i) amino acid sequences having at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 1, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 1; and preferably such amino acid sequences that comprise the amino acid residues H32, A33, H35, Y50, and S99, with respect to SEQ ID NO.
  • amino acid sequences will, as mentioned herein, provide recognition support for mono-Tn epitopes
  • amino acid sequences having at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 28, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 28; and preferably such amino acid sequences that comprise the amino acid residues T28, T30, H32, A33, H35, Y50, S99, L101, A102 and L103, with respect to SEQ ID NO. 28 (which amino acid sequences will, as mentioned herein, provide recognition support for mono- STn epitopes).
  • VH sequences libraries of the invention comprising and/or based on such VH sequences, methods for generating such libraries and antibodies identified using, generated using and/or isolated from such libraries form further preferred aspects of the invention.
  • the invention provides a method for generating antibodies against a target that is present on the surface of a cancer cell (and/or is expressed by a cancer cell on its surface), in which said target is present on/expressed on the cancer cell in a glycosylated form (and in particular in a glycosylated form that contains, comprises or presents one or more Tn ot STn epitopes, as as described herein), which method at least comprises the following steps:
  • each antibody in the library comprises a VH domain and VL domain, in which each such VH domain is (i) an amino acid sequence having at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 1, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 1; and preferably such an amino acid sequence that comprises the amino acid residues H32, A33, H35, Y50, and S99, with respect to SEQ ID NO.
  • amino acid sequence having at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 28, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 28; and preferably such an amino acid sequences that comprises the amino acid residues T28, T30, H32, A33, H35, Y50, S99, L101, A102 and L103, with respect to SEQ ID NO.
  • the library is preferably such that it comprises (a collection or repertoire of) at least 1000 (10 3 ), such as at least 10.000 (10 4 ), preferably at least 100.000 (10 5 ), such as 10 5 or more, different VL domains such that, upon each of said VL domains combining with said VH domain as part of an antibody that can be obtained from said library, said library comprises/provides (a collection or repertoire of) at least 1000 (10 3 ), such as at least 10.000 (10 4 ), preferably at least 100.000 (10 5 ), such as 10 5 or more, different antibodies (i.e.
  • antibodies having different VH/VL combinations screening said library against said target, and preferably screening said library against said target in which said target is in a glycosylated form, and even more preferably screening said library against said target in which said target is in a glycosylated form that contains, comprises or presents one or more Tn or STn epitopes; identifying, generating and/or isolating one or more antibodies (or sequences encoding such antibodies) that are capable of specifically binding to said target, and that are preferably capable of specifically binding to said target in a glycosylated form, and more preferably capable of specifically binding to said target in a glycosylated form that contains, comprises or presents one or more Tn or STn epitopes; and optionally applying one or more steps for improving one or desired properties of the antibody/antibodies thus obtained, for example by means of humanization, affinity maturation, removing potential immunogenic epitopes and/or optimizing the sequence for expression or production in a desired host cell or host organism.
  • the invention provides a method for generating antibodies against a target that is present on the surface of a cancer cell (and/or is expressed by a cancer cell on its surface), in which said target is present on/expressed on the cancer cell in a glycosylated form (and in particular in a glycosylated form that contains, comprises or presents one or more Tn ot STn epitopes, as as described herein), which method at least comprises the following steps:
  • each antibody in the library comprises a VH domain and VL domain, in which each such VH domain has the amino acid sequence of SEQ ID NO. 1 and/or the amino acid sequence of SEQ ID NO.
  • the library is preferably such that it comprises (a collection or repertoire of) at least 1000 (10 3 ), such as at least 10.000 (10 4 ), preferably at least 100.000 (10 5 ), such as 10 5 or more, different VL domains such that, upon each of said VL domains combining with said VH domain as part of an antibody that can be obtained from said library, said library comprises/provides (a collection or repertoire of) at least 1000 (10 3 ), such as at least 10.000 (10 4 ), preferably at least 100.000 (IO 5 ), such as IO 5 or more, different antibodies (i.e.
  • antibodies having different VH/VL combinations screening said library against said target, and preferably screening said library against said target in which said target is in a glycosylated form, and even more preferably screening said library against said target in which said target is in a glycosylated form that contains, comprises or presents one or more Tn or STn epitopes; identifying, generating and/or isolating one or more antibodies (or sequences encoding such antibodies) that are capable of specifically binding to said target, and that are preferably capable of specifically binding to said target in a glycosylated form, and more preferably capable of specifically binding to said target in a glycosylated form that contains, comprises or presents one or more Tn or STn epitopes; and optionally applying one or more steps for improving one or desired properties of the antibody/antibodies thus obtained, for example by means of humanization, affinity maturation, removing potential immunogenic epitopes and/or optimizing the sequence for expression or production in a desired host cell or host organism.
  • the invention provides a method for generating antibodies against a target that is present on the surface of a cancer cell (and/or is expressed by a cancer cell on its surface), in which said target is present on/expressed on the cancer cell in a glycosylated form (and in particular in a glycosylated form that contains, comprises or presents one or more Tn ot STn epitopes, as as described herein), which method at least comprises the following steps:
  • each antibody in the library comprises a VH domain and VL domain
  • each such VH domain comprises: (i) a CDR1 having the amino acid sequence DHAIH (SEQ ID NO: 155) or an amino acid sequence that has a single amino acid difference (as defined herein) with the amino acid sequence DHAIH (SEQ ID NO: 155); and (ii) a CDR2 having the amino acid sequence YISPGNDDIKYNEKFKG (SEQ ID NO: 156) or an amino acid sequence that has 3, 2 or 1 (and preferably 2 or 1, and most preferably 1) amino acid differences (as defined herein) with the amino acid sequence YISPGNDDIKYNEKFKG (SEQ ID NO: 156); and (iii) a CDR3 having the amino acid sequence SLPGTFDY (SEQ ID NO: 157) or an amino acid sequence that has a single amino acid difference (as defined here
  • antibodies having different VH/VL combinations screening said library against said target, and preferably screening said library against said target in which said target is in a glycosylated form, and even more preferably screening said library against said target in which said target is in a glycosylated form that contains, comprises or presents one or more Tn or STn epitopes; identifying, generating and/or isolating one or more antibodies (or sequences encoding such antibodies) that are capable of specifically binding to said target, and that are preferably capable of specifically binding to said target in a glycosylated form, and more preferably capable of specifically binding to said target in a glycosylated form that contains, comprises or presents one or more Tn or STn epitopes; and optionally applying one or more steps for improving one or desired properties of the antibody/antibodies thus obtained, for example by means of humanization, affinity maturation, removing potential immunogenic epitopes and/or optimizing the sequence for expression or production in a desired host cell or host organism.
  • a VH sequence when a VH sequence comprises a CDR1 that is an amino acid sequence that has a single amino acid difference (as defined herein) with the amino acid sequence DHAIH (SEQ ID NO: 155), then such a CDR1 preferably comprises the amino acid residues H32, A33, and H35; when a VH sequence comprises a CDR2 that is an amino acid sequence that has 3, 2 or 1 (and preferably 2 or 1, and most preferably 1) amino acid differences (as defined herein) with the amino acid sequence YISPGNDDIKYNEKFKG (SEQ ID NO: 156), then such a CDR2 preferably comprises the amino acid residues Y5O, S52, N55, and D57; and when a VH sequence comprises a CDR3 that is an amino acid sequence that has a single amino acid difference (as defined herein) with the amino acid sequence SLPGTFDY (SEQ ID NO: 157), then such a CDR3 preferably comprises
  • a library that both comprises one or more VH sequences that have been chosen to confer specificity for a Tn epitope (or a protein that contains, comprises or presents a Tn epitope) and also comprises one or more VH sequences that have been chosen to confer specificity for an STn epitope (or a protein that contains, comprises or presents a STn epitope).
  • such a library can be suitably constructed using (or suitably comprise) both the VH sequence of SEQ ID NO: 1 (or a VH sequence based on the VH sequence of SEQ ID NO: 1, as further described herein) and the VH sequence of SEQ ID NO: 28 (or a VH sequence based on the VH sequence of SEQ ID NO: 28, as further described herein).
  • Such a library can for example be used to generate and/or screen for antibodies that are expressed on a cancer cell, irrespective of whether said protein comprises or presents a Tn epitope, an STn epitope, or both.
  • one aspect of the invention relates to a library as described herein that contains at least one VH sequence that confers, to the antibodies that can be obtained from such a library, specificity for a Tn epitope and further contains at least one VH sequence that confers, to the antibodies that can be obtained from such a library, specificity for a STn epitope.
  • the VH domain conferring specificity to the Tn epitope may be the VH domain of SEQ ID NO: 1 (or a variant thereof as described herein and/or a VH domain having CDRs that are based on the CDRs of SEQ ID NO: 1, as also described herein) and the VH domain conferring specificity to the STn epitope may be the VH domain of SEQ ID NO: 28 (or a variant thereof as described herein and/or a VH domain having CDRs that are based on the CDRs of SEQ ID NO:28, as also described herein)
  • the methods of the invention will usually involve the use of a library that only contains one or more VH sequence(s) that are capable of specifically binding to a Tn epitope (and/or have been chosen based on their ability to specifical bind to a Tn epitope).
  • the VH domain conferring specificity to the Tn epitope may be the VH domain of SEQ ID NO: 1 (or a variant thereof as described herein and/or a VH domain having CDRs that are based on the CDRs of SEQ ID NO: 1, as also described herein); and libraries that only contain VH sequences that are capable of specifically binding to a Tn epitope (and/or have been chosen based on their ability to specifical bind to a Tn epitope) form further aspects of the invention.
  • the methods of the invention will usually involve the use of a library that only contains one or more VH sequence(s) that are capable of specifically binding to a STn epitope (and/or have been chosen based on their ability to specifical bind to a STn epitope).
  • the VH domain conferring specificity to the STn epitope may be the VH domain of SEQ ID NO: 28 (or a variant thereof as described herein and/or a VH domain having CDRs that are based on the CDRs of SEQ ID NO:28, as also described herein); and libraries that only contain VH sequences that are capable of specifically binding to a STn epitope (and/or have been chosen based on their ability to specifical bind to a STn epitope) form further aspects of the invention.
  • the (collection or repertoire of different) VL sequences that can be present/included in the library may be provided and/or have been generated in any suitable manner known per se, and may for example be a collection or repertoire of naive VL sequences (for example derived from a naive library of antibody sequences obtained from mouse B-cells or human B-cells), pre- immume VL sequences, synthetic VL sequences and/or semi-synthetic VL sequences; or any combination of the foregoing.
  • naive VL sequences for example derived from a naive library of antibody sequences obtained from mouse B-cells or human B-cells
  • pre- immume VL sequences for example derived from a naive library of antibody sequences obtained from mouse B-cells or human B-cells
  • synthetic VL sequences and/or semi-synthetic VL sequences
  • these and other suitable techniques may be used or suitably adapted to provide a library in which the desired VL sequences are suitably combined with one or more VH sequences that have been chosen to specifically bind to a Tn or STn epitope, respectively; so as to provide a library of the invention as further described herein, in which the size and/or diversity of the library is mainly provided by (the collection or repertoire of) the different VL sequences that are present in the library.
  • the libraries of the invention may contain or comprise a collection or repertoire of naive VL sequences, a collection or repertoire of synthetic VL sequences, or a collection or repertoire of semi-synthetic VL sequences (with libraries based on a collection or repertoire of naive VL sequences forming one preferred aspect of the invention), in each case combined with a VH sequence that is as further described herein.
  • an immune repertoire of VL sequences is not excluded from the invention in its broadest sense, the use of an immune repertoire will usually be less preferred, as the VL sequences derived from an immune VH/VL repertoire may often requiring pairing with the specific VH sequence that they are associated with in the immune repertoire in order to provide the degree of specificity that is intended for the purposes of the present invention.
  • the VL domain in the antibody is preferably such that it is capable of binding to (part of) the peptide backbone that, in the intended or desired target, is associated with the Tn epitope to which VH sequence in the antibody can bind.
  • the antibodies that can be obtained from each of these libraries i.e.
  • the VL domain in the antibody is preferably such that it is capable of binding to (part of) the peptide backbone that, in the intended or desired target, is associated with the STn epitope to which VH sequence in the antibody can bind.
  • (the sequence of) the VH domain is preferably such that it essentially does not contribute to the binding of the antibody to the peptide epitope (i.e. to the associated part of the peptide backbone of the target).
  • the methods and libraries provided by the invention can in particular be used to obtain, identify and/or generate (sequences encoding) antibodies against proteins/targets that are present on/expressed on cancer cells, in particular against proteins/targets that are present on/expressed on cancer cells in a glycosylated form, and more in particular against proteins/targets that are present on/expressed on cancer cells in a glycosylated form where said glycosylated form comprises, contains or presents one or more Tn and/or STn epitopes (as further described herein).
  • proteins/targets as well as the types of cancer cells on which they are expressed (and sometimes overexpressed compared to healthy cells and/or other cancer cells) and the cancer types with which such proteins, targets and cells are associated will be clear to the skilled person, and for example and without limitation include: EGFR, VEGFR, HER2, CD37, FLT3, FGFR, CD19, CD22, CD27, CD25, CD30, CD33, CD38, CD43, Mesothelin, PD-L1, CD44, Podocalyxin (TRA1.60/80), CD133, CD90, CD326, Cripto-1, ABCG2, CD24, CD49, Notch2, CD146 (MUC18), CD10, CD117, CD26, CXCR4, CD34, CD271, CD13, CD56, CD105, LGR5, CD114, CD54, CXCR1, TIM-3, CD55, DLL-4, CD96, CD29, CD9, CD166, CD44, ABCB5, Notch3, CD123, MUC1, MUC4, M
  • an antibody for treating a specific type of cancer when it is desired to generate an antibody for treating a specific type of cancer, a person skilled in the field of oncology will be able to select a protein or target that will be associated with cancer cells that are involved in said type of cancer and then use antibodies against said target(s) that have been generated using the methods and libraries described herein for treating said type of cancer.
  • the skilled person will be able to determine which (glycosylated) proteins are overexpressed on the (type of) cancer cell involved and then use the methods and libraries of the invention to generate one or more antibodies against saud proteins.
  • the latter aspect of the invention may find use in the area of so-called "personalized medicine", by allowing the skilled person to generate antibodies that are directed against proteins that are expressed on cancer cells that have been obtained from the patient to be treated, and then use said antibodies to treat said patient.
  • the methods and libraries of the invention are used to generate antibodies against a protein/target belonging to the MUC family that are expressed (in a glycosylated form) on a cancer cell.
  • the invention also relates to methods for treating cancer which comprise administering, to a patient in need thereof, one or more a therapeutically effective amounts of an antibody that can be obtained or have been obtained from the libraries and/or using the methods described herein.
  • the invention further relates to methods for treating cancer which comprise administering, to a patient suffering from cancer, one or more a therapeutically effective amounts of an antibody that can be obtained or have been obtained from the libraries and/or using the methods described herein, in which said antibody is directed against a (glycosylated) protein that is present on/expressed by a cancer cell that is present in the body of the patient to be treated.
  • the invention also relates antibodies that can be obtained or have been obtained from the libraries and/or using the methods described herein for the treatment of cancer.
  • the antibodies used can be as further described herein.
  • such antibodies are directed against a protein/target belonging to the MUC family (as further described herein).
  • Tn antigen expression is mainly observed in breast, colorectal and pancreatic tumors while STn antigen was highly expressed in colorectal and pancreatic tumors.
  • a further aspect of the invention relates to the use of antibody that can be obtained or have been obtained from the libraries and/or using the methods described herein for the treatment of cancer, where the cells of the tumor involved and/or the type of cancer involved mainly express proteins on their surface that contain or present Tn antigens, in which the antibody used comprises a VH domain that confers specificity for a Tn antigen and/or wherein the antibody has been obtained from a library as described herein that was constructed using VH sequences that confer specificity for a Tn antigen.
  • VL sequence present in such an antibody will be as further described herein and may in particular confer specificity for (the protein backbone of) a protein present or expressed on the cancer cell that contains or presents the Tn antigen.
  • Another aspect of the invention relates to the use of antibody that can be obtained or have been obtained from the libraries and/or using the methods described herein for the treatment of cancer, where the cells of the tumor involved and/or the type of cancer involved mainly express proteins on their surface that contain or present STn antigens, in which the antibody used comprises a VH domain that confers specificity for a STn antigen and/or wherein the antibody has been obtained from a library as described herein that was constructed using VH sequences that confer specificity for a STn antigen.
  • VL sequence present in such an antibody will be as further described herein and may in particular confer specificity for (the protein backbone of) a protein present or expressed on the cancer cell that contains or presents the STn antigen.
  • the invention relates to antibodies that can be obtained or have been obtained from the libraries described herein and/or using the methods described herein.
  • such antibodies contain, as a VH domain, an amino acid sequence that is the VH sequence of SEQ ID NO: 1, or a variant of the VH sequence of SEQ ID NO: 1 (which variant is preferably as further described herein) or a VH sequence that has CDRs that are the same as the CDR sequences that are present in the VH sequence of SEQ ID NO: 1 or that have been derived from the CDR sequences that are present in the VH sequence of SEQ ID NO: 1.
  • Said CDR sequences present in the VH sequence of SEQ ID NO: 1 are as follows:
  • CDR1 DHAIH (SEQ ID NO: 155);
  • CDR2 YISPGNDDIKYNEKFKG (SEQ ID NO: 156);
  • CDR3 SLPGTFDY (SEQ ID NO: 157).
  • such antibodies contain, as a VH domain, an amino acid sequence that is the VH sequence of SEQ ID NO: 28, or a variant of the VH sequence of SEQ ID NO: 28 (which variant is preferably as further described herein) or a VH sequence that has CDRs that are the same as the CDR sequences that are present in the VH sequence of SEQ ID NO: 28 or that have been derived from the CDR sequences that are present in the VH sequence of SEQ ID NO: 28.
  • Said CDR sequences present in the VH sequence of SEQ ID NO: 28 are as follows:
  • CDR1 DHAIH (SEQ ID NO: 155);
  • CDR2 YISPGNDDIKYNEKFKG (SEQ ID NO: 156);
  • the invention relates to an antibody, and in particular an antibody that can be obtained or has been obtained using the methods described herein and/or from a library as described herein, which antibody comprises a VH domain and a VL domain, in which the VH domain comprises: a CDR1 having the amino acid sequence DHAIH (SEQ ID NO: 155) or an amino acid sequence that has a single amino acid difference (as defined herein) with the amino acid sequence DHAIH (SEQ ID NO: 155); and
  • a CDR2 having the amino acid sequence YISPGNDDIKYNEKFKG (SEQ ID NO: 156) or an amino acid sequence that has 3, 2 or 1 (and preferably 2 or 1, and most preferably 1) amino acid differences (as defined herein) with the amino acid sequence YISPGNDDIKYNEKFKG (SEQ ID NO: 156); and a CDR3 having the amino acid sequence SLPGTFDY (SEQ ID NO: 157) or an amino acid sequence that has a single amino acid difference (as defined herein) with the amino acid sequence SLPGTFDY (SEQ ID NO: 157); and in which the VL domain is as further described herein, and is in particular such that said VL domain is capable of binding to the peptide backbone of a glycosylated protein that is present or expressed on the surface of a cancer cell (and more in particular capable of binding to the peptide backbone of a glycosylated protein that is present or expressed on the surface of a cancer cell, where said glycos
  • the VH sequence that is present in the antibody comprises a CDR1 that is an amino acid sequence that has a single amino acid difference (as defined herein) with the amino acid sequence DHAIH (SEQ ID NO: 155)
  • such a CDR1 preferably comprises the amino acid residues H32, A33, and H35
  • the VH sequence that is present in the antibody a CDR2 that is an amino acid sequence that has 3, 2 or 1 (and preferably 2 or 1, and most preferably 1) amino acid differences (as defined herein) with the amino acid sequence YISPGNDDIKYNEKFKG (SEQ ID NO: 156)
  • such a CDR2 preferably comprises the amino acid residues Y50, S52, N55, and D57
  • the VH sequence that is present in the antibody comprises a CDR3 that is an amino acid sequence that has a single amino acid difference (as defined herein) with the amino
  • the invention relates to an antibody, and in particular an antibody that can be obtained or has been obtained using the methods described herein and/or from a library as described herein, which antibody comprises a VH domain and a VL domain, in which the VH domain comprises a CDR1 having the amino acid sequence DHAIH and a CDR2 having the amino acid sequence YISPGNDDIKYNEKFKG and a CDR3 having the amino acid sequence SLPGTFDY (SEQ ID NO: 157); and in which the VL domain is as further described herein, and is in particular such that said VL domain is capable of binding to the peptide backbone of a glycosylated protein that is present or expressed on the surface of a cancer cell (and more in particular capable of binding to the peptide backbone of a glycosylated protein that is present or expressed on the surface of a cancer cell, where said glycosylated protein contains or presents a Tn antigen) and/or such that said VL
  • the invention relates to an antibody, and in particular an antibody that can be obtained or has been obtained using the methods described herein and/or from a library as described herein, which antibody comprises a VH domain and a VL domain, in which the VH domain comprises: a CDR1 having the amino acid sequence DHAIH (SEQ ID NO: 155) or an amino acid sequence that has a single amino acid difference (as defined herein) with the amino acid sequence DHAIH (SEQ ID NO: 155); and
  • a CDR2 having the amino acid sequence YISPGNDDIKYNEKFKG (SEQ ID NO: 156) or an amino acid sequence that has 3, 2 or 1 (and preferably 2 or 1, and most preferably 1) amino acid differences (as defined herein) with the amino acid sequence YISPGNDDIKYNEKFKG (SEQ ID NO: 156); and a CDR3 having the amino acid sequence SLLALDY (SEQ ID NO: 158) or an amino acid sequence that has a single amino acid difference (as defined herein) with the amino acid sequence SLLALDY (SEQ ID NO: 158) and in which the VL domain is as further described herein, and is in particular such that said VL domain is capable of binding to the peptide backbone of a glycosylated protein that is present or expressed on the surface of a cancer cell (and more in particular capable of binding to the peptide backbone of a glycosylated protein that is present or expressed on the surface of a cancer cell, where said glycosylated
  • the VH sequence that is present in the antibody comprises a CDR1 that is an amino acid sequence that has a single amino acid difference (as defined herein) with the amino acid sequence DHAIH (SEQ ID NO: 155)
  • such a CDR1 preferably comprises the amino acid residues H32, A33, and H35
  • the VH sequence that is present in the antibody a CDR2 that is an amino acid sequence that has 3, 2 or 1 (and preferably 2 or 1, and most preferably 1) amino acid differences (as defined herein) with the amino acid sequence YISPGNDDIKYNEKFKG (SEQ ID NO: 156)
  • such a CDR2 preferably comprises the amino acid residues Y50, S52, N55, and D57
  • the VH sequence that is present in the antibody comprises a CDR3 that is an amino acid sequence that has a single amino acid difference (as defined herein) with the amino
  • the invention relates to an antibody, and in particular an antibody that can be obtained or has been obtained using the methods described herein and/or from a library as described herein, which antibody comprises a VH domain and a VL domain, in which the VH domain comprises a CDR1 having the amino acid sequence DHAIH (SEQ ID NO: 155) and a CDR2 having the amino acid sequence YISPGNDDIKYNEKFKG (SEQ ID NO: 156) and a CDR3 having the amino acid sequence SLLALDY (SEQ ID NO: 158); and in which the VL domain is as further described herein, and is in particular such that said VL domain is capable of binding to the peptide backbone of a glycosylated protein that is present or expressed on the surface of a cancer cell (and more in particular capable of binding to the peptide backbone of a glycosylated protein that is present or expressed on the surface of a cancer cell, where said glycosylated protein contains or presents a STn
  • antibodies according to the preceding aspects can be as further described herein.
  • further aspects of the invention relate to nucleotide sequences that encode an antibody according to one of the preceding aspects, to host cells and/or host organisms that express or can be used to produce an antibody according to one of the preceding aspects, to pharmaceutical compositions that contain at least one antibody according to one of the preceding aspects, and to uses of an antibody according to one of the preceding aspects; all of which are preferably as further described herein.
  • nucleic acid encompasses double as well as single-stranded nucleotide molecules. Nucleic acid sequences, when provided, are listed in the 5' to 3' direction, unless stated otherwise.
  • homology is determined by comparing the amino acid sequence and its conserved amino acid substitutes of one protein sequence to the second protein sequence. Similarity may be determined by procedures which are well-known in the art, for example, a BLAST program (Basic Local Alignment Search Tool at the National Center for Biological Information). Likewise, homology, similarity or sequence identity between two nucleic acid sequences may be determined by procedures which are well-known in the art, for example, a BLAST program (Basic Local Alignment Search Tool at the National Center for Biological Information).
  • amino acid residue substitution at a specific position means substitution with any amino acid different from the native amino acid residue that is present at that specific position.
  • a conservative amino acid substitution replaces an amino acid with another amino acid that is similar in size and chemical properties such that the substitution has no or only minor effect on protein structure and function; meanshile a nonconservative amino acid substitution replaces an amino acid with another amino acid that is dissimilar and thereby is likely to affect structure and function of the protein.
  • antibody will be understood to include proteins having the characteristic two-armed, Y-shape of a typical antibody molecule as well as one or more fragments of an antibody that retain the ability to specifically bind to an antigen.
  • exemplary antibodies include, but are not limited to, a monoclonal antibody, a polyclonal antibody, a bi-specific antibody, a multispecific antibody, a grafted antibody, a human antibody, a humanized antibody, a synthetic antibody, a chimeric antibody, a camelized antibody, a single-chain Fvs (scFv) (including fragments in which the VL and VH are joined using recombinant methods by a synthetic or natural linker that enables them to be made as a single protein chain in which the VL and VH regions pair to form monovalent molecules, including single chain Fab and scFab), a single chain antibody, a Fab fragment (including monovalent fragments comprising the VL, VH, CL, and CHI domains),
  • scFv
  • VL refers to antibody variable domain, light chain
  • VH refers to antibody variable domain, heavy chain
  • Tn antigen refers to the monosaccharide structure N- acetylgalactosamine (GalNAc) linked to serine (Ser) or threonine (Thr) on a peptide backbone by a glycosidic bond (i.e. GalNAcal-O-Ser/Thr). The initials stand for Thomsen-nouveau. Tn antigen is expressed in most carcinomas.
  • Tn- carbohydrate epitope (or “Tn epitope”) as used herein refers to the GalNac part of the Tn antigen.
  • mono-Sn refers to one Tn moiety (i.e. one GalNAc).
  • bis-Tn refers to two Tn moieties (i.e. two GalNAc).
  • STn antigen refers to a sialyl-Tn antigen, formed by elongation of the Tn antigen with sialic acid (Neu5Ac(a2-6)GalNAc), still linked to serine (Ser) or threonine (Thr) (i.e. Neu5Aca2-6GalNAcal-O-Ser/Thr).
  • Tn and STn may have additional modifications, such as phosphorylation, acetylation, methylation, and sulfonation.
  • STn- carbohydrate epitope refers to the Neu5Ac(a2- 6)GalNAc part of the Tn antigen.
  • mono-STn refers to one STn moiety (i.e. one Neu5Ac(a2-6)GalNAc).
  • bis-STn refers to two STn moieties (i.e. two Neu5Ac(a2-6)GalNAc).
  • glycoprotein generally refers to proteins which contain oligosaccharide chains covalently attached to amino acid side-chains.
  • glycoprotein refers to a protein or a peptide thereof which contains a carbohydrate moiety (preferably a Tn or STn epitope) covalently attached to an amino acid residue (such as serine, threonine, or tyrosine; or any non-natural amino acid derivatives thereof such as replacing O with S) of said protein or peptide thereof.
  • combotope refers to the combination of a carbohydrate epitope and a peptide epitope being recognized by an antibody.
  • the two epitopes form a common epitope, the "combotope", which is different from each of the two epitopes from which it is composed.
  • the peptide epitope is associated with a Tn epitope or a STn epitope; the peptide epitope may be associated with one or two Tn or STn epitopes.
  • the combotope may be contiguous or discontiguous - i.e.
  • the carbohydrate epitope is directly attached to the petide epitope by covalent bond (contiguous), or the carbohydrate epitope is attached to an amino acid residue located 1, 2, 3, 4, 5, 6, 7, 8, 9, or 10 amino acid residues up or down stream of the peptide epitope (discontiguous).
  • the carbohydrate epitope is covalently attached to the peptipe epitope.
  • the carbohydrate epitope may also be a combination of two Tn or STn epitopes (termed bis-Tn and bis-STn) or one Tn and one STn epitope, on adjacent amino acids in said peptide sequence in said glycoprotein. Adjacent means separated by 0 to 2 amino acids.
  • Antibodies which are refered to as “combotope binder” recognize and bind the combination of both the carbohydrate epitope and a peptide epitope of the glycoprotein. Such antibodies are also referred to as “combotope antibodies”.
  • hapten refers to the carbohydrate moiety(ies) (Tn or STn) independent of the peptide/protein.
  • Antibodies which are referred to as “hapten binders” will bind Tn or STn epitopes independent of the peptide/protein carries - hence, they are therefore not specific for the combination of both the carbohydrate epitope and a peptide epitope of the glycoprotein (i.e. not a combotope binder).
  • phase "domain which binds" as used herein should be understood as “domain which is suitable for binding", “domain which is capable of binding", and/or “domain which is prepared for binding".
  • FIG. 1 Schematic illustration of an embodiment of the invention: Tn-template antibody library for identifying Tn-combotope antiobodies. Phage display of a library of scFv antibodes. Each scFv in the library has a Tn-binding VH domain. The library is screened for Tn-peptide specific scFv by biopanning using Tn-peptides.
  • Figure 2 Illustration of the prepararion of a phage display library of the present invention. 1) mRNA isolation from mouse spleen. 2) cDNA synthesis with reverse transcriptase using random hexamers. 3) PCR amplification from cDNA template to obtain the VH domain using a specific set primers, and the repertoire of VL domains using a mix of VL primesr. 4) PCR assembly of VL-domain repertoire and specific VH- domain using 5' phosphorylated outer primers.
  • Figure 3 X-ray of G2D11 scFv with APGS*T*AP peptide (where * denotes a GalNac residue) showed the interaction points of VH with the glycan structure. Key interaction points with two adjacent GalNac residues unclude His32 H , Ala33 H , His35 H in CDR1; His40 H , Ser52 H , Asn55 H , Asp57 H in CDR2; and Ser99 H in CRD3.
  • FIG. 4 VH domain sequence alignment of G2D11 with other known VH-domains. conserveed amino acid residues are indicated by arrows (H32, A33, H35, Y50, S52, N55, D57 AND S99)
  • FIG. 5 Phage and sequence enrichment after each round (Round 1, 2, and 3) of biopanning for bisTn-MUCl.
  • Polyclonal phage ELISA confirmed phage enrichments for bis Tn-MUCl target peptide.
  • bisTn MUC1 peptide no 1 in Table 1;
  • Tn MUC1 peptide no 3 in Table 1;
  • SA negative control.
  • FIG. 1 MUC1 monoclonal ELISA. Screening of monoclonal scFvs on MUC1 target and control peptides.
  • FIG. 8 MUC1 scFvs binding assays and kinetic affinities.
  • A scFv titration at fixed concentration of MUC1 target peptide (peptide no 1 in Table 1).
  • B scFv titration at fixed concentration of IgA hinge region control peptide (peptide no 8). Each data point is the mean value of three independent experiments.
  • C Representative histograms of cell binding at 1.25 pg/ml of A3, D2 and D3 scFvs along with 5E5 mAb on MDA-MB-231 WT and COSMC KO cells. Flow cytometry experiments were repeated three times.
  • FIG. 10 MUC1 scFvs biological evaluation with flow cytometry.
  • A Negative binding of MUC1 scFvs on HEK293 cells as a negative control cell line.
  • B MCF7 cells at 1.25 pg/mL.
  • C Representative example of concentration dependent binding on MDA-MB-231 WT and COSMC KO cells.
  • scFv D3 is shown at 4-fold dilution starting from 5 pg/mL.
  • FIG. 12 Heat map of binding of scFv D3, scFv A4, scFv 5E5, scFv 2D9Chi, and scFv G2D11 to Tn-glycopeptides from Table 5.
  • the glycopeptides were printed on a microarray chip.
  • the heat map shows amino acids 9-19 of the peptides in Table 5.
  • the relative fluorescence units (RFU) as shown as heat map. Tn-glycosylation sites are bold and underlined. Substitutions with Ala are marked as bold.
  • FIG. 14 CD43 monoclonal ELISA. Screening of monoclonal scFvs on CD43 target and control peptides.
  • FIG. 15 CD43 scFvs binding assays.
  • A Eight scFvs were were titrated on bisTn- CD43 target peptide (peptide no. 9 in Table 1).
  • B ScFv titration on IgAl hinge region control glycopeptide (peptide no. 8 in Table 1) showed A7, D3 cross reactivity to IgA while Al and F4 showed weaker binding to IgAl. Each dot represents the mean value of three independent experiments.
  • C Representative histograms of Al, D7, Hl and H2 scFvs at 1.25 pg/mL tested on Jurkat cells before and after neuraminidase treatment. Flow cytometry experiments were repeated three times.
  • FIG. 16 CD43 scFvs biological evaluation with flow cytometry. Concentration dependent binding of Al scFv as a representative example on HEK293 cells and Jurkat cells, before and after neuraminidase treatment. scFv was 4-fold diluted starting from 5 pg/mL.
  • Figure 17. CD43 x-ray structure with GAS*T*GSP peptide reveales the importance of Tyr99L as key interaction point with the peptide backbone.
  • FIG. 1 Alignment of bisTn binder G2D11 VH with monoTn binder 3F1 VH to identify amino acid residues relevant for shifting to anti bisSTn.
  • Microaray data for binding of G2D11, 3F1, and mutants (M l-4) comprising selected mutations of VH-G2D11 to glycopeptides 1 (bisTnMUCl), 11 (bisSTnMUCl), 12 (monoSTnMUCl) and 4 (unglycosylated control).
  • FIG. 20 Microaray data for binding of STnMUCl-D4, D3, C7 scFv to glycopeptides 1 (bisTnMUCl), 11 (bisSTnMUCl), and 12 (monoSTnMUCl).
  • FIG. 21 ELISA titration screening of the humanized scFvs on coated Tn-MUCl and other Tn-proteins.
  • A D3LlHlscFv
  • B D3LlH2scFv
  • C D3L2H3scFv
  • D D3L3H4scFv
  • E D3L4H5scFv
  • F parental mouse D3.
  • MUCl mucin 1
  • MUC21 mucin 21
  • GPNMB Transmembrane Glycoprotein NMB
  • EGFR Epidermal growth factor receptor
  • VVL Vicia Villosa Lectin used for detecting Tn on the Tn-proteins.
  • Figure 22 Illustration of (A) bisTnMUCl, (B) monoTn(Thr)MUCl, and (C) monoTn(Ser)MUCl.
  • Figure 23 Elisa titration of different target Tn-peptides detected with mouse D3 (lug/mL) scFvs on streptavidin coated plates.
  • Column 1 Biotin-2OEG-2OEG- HSSSTIPTPA(MUC13)
  • column 2 Biotin-2OEG-2OEG-HSSSTIPIPT(MUC13)
  • column 3 Biotin-2OEG-2OEG-SESITNVNSL(MUC13)
  • column 4 Biotin-2OEG-2OEG-
  • the present invention concerns a new antibody concept technology for simple and rapid development of antibodies targeting Tn- and STn- glycosylation sites of any glycoprotein site of choice.
  • the invention concerns antibody libraries which can be screened for antibodies which have improved specificity due to their specificity towards a combination of an epitopes on a carbohydrate part of a glycoprotein and an epitope on a peptide backbone in said glycoprotein which is associated with the carbohydrate epitope.
  • the combined epitope is termed a "combotope”.
  • a non-limiting embodiment of the invention is schematically illustrated in Figure 1.
  • the present invention is especially useful in the generation of therapeutics for cancer treatment and diagnostics.
  • the present invention provides combotope antibodies which have high specificity and high binding efficiency to their target glycopeptide due to their combined specificity towards both the carbohydrate epitope as well as the peptide backbone epitope associated with the carbohydrate epitope of the glycoprotein target.
  • the present invention provides a combotope antibody for targeting tumor cells carrying said glycoprotein on their surface.
  • the invention provides an antibody for targeting tumor cells, said antibody comprising two antibody domains, where the first antibody domain binds a carbohydrate epitope of a glycoprotein of the tumor cell and the second antibody domain binds a peptide epitope of said glycoprotein of said tumor cell, where said antibody specifically binds both epitopes (as a common epitope) as compared to only binding one of said epitopes.
  • the first antibody domain is a VH domain and the second antibody domain is a VL domain. In another embodiment, both the first and second antibody domains are VH domains, but different from one another.
  • the antibodies disclosed herein are scFv antibodies, comprising a VH domain and a VL domain, where both domains are present in a single polypeptide chain.
  • the Fv polypeptide further comprises a polypeptide linker between the VH and VL domains allowing the scFv to form the desired structure for antigen binding.
  • the linker is selected from (GGGGS) n where in is 1, 2, 3, 4, 5, or 6; e.g. (GGGGS)4 or (GGGGS)s.
  • Other linkers could be used as an alternative to the exemplified (GGGGS) n -linker. The skilled person would know how to select such linkers.
  • the invention provides an antibody for targeting tumor cells, said antibody comprising (i) a VH domain binding a carbohydrate epitope of a glycoprotein of the tumor cell and (ii) a VL domain binding a peptide epitope of said glycoprotein of said tumor cell, where said antibody is selected to specifically binding both epitopes (as a common epitope) as compared to only binding one of said epitopes.
  • the invention provides an antibody for targeting a glycoprotein, such as a specific glycoprotein on a tumor cell.
  • the glycoprotein comprises a carbohydrate epitope and a peptide epitope.
  • the carbohydrate epitope is a short truncated O-glycan, such as Tn or STn.
  • the peptide epitope is associated with the carbohydrate epitope, such as directly attached by virtue of being chemically linked or being in close proximity of one another by virtue of the structural configuration of the glycoprotein.
  • the antibody of the present invention is characterized by its ability to specifically bind both epitopes (as a common epitope), as compared to only binding one of said epitopes.
  • the antibody comprises (i) a VH domain characterized by (a) being suitable for binding a carbohydrate epitope of a glycoprotein of a tumor cell and (b) not binding the peptide epitope of said glycoprotein of said tumor cell and (ii) a VL domain characterized by being suitable for binding the peptide epitope of said glycoprotein of said tumor cell.
  • the VL-domain is further characterized by not binding the carbohydrate epitope of said glycoprotein.
  • the antibody only binds a combination of carbohydrate and peptide epitopes, i.e. both epitops need to be present for binding of the antibody.
  • the present invention provides an antibody which binds a tumor cell, wherein the antibody comprises
  • the present invention provides an antibody which binds a tumor cell, wherein the antibody comprises (ii) a VL-domain which binds a peptide epitope of a glycoprotein of said tumor cell and (i) a VH-domain which binds a carbohydrate epitope of said glycoprotein of said tumor cell and which does not contribute to or interfere with binding said peptide epitope of said glycoprotein of said tumor cell, i.e.
  • the VH-domain binding is not affected/influenced by the presense of any peptide epitope; and wherein the peptide epitope of the glycoprotein of the tumor cell is associated with the carbohydrate epitope of said glycoprotein of said tumor cells, and wherein the antibody is specific for the combination of said carbohydrate epitope and said peptide epitope on said glycoprotein of said cancer cell.
  • the present invention provides an antibody which binds a tumor cell, wherein the antibody comprises (ii) a VL-domain which binds a peptide epitope of a glycoprotein of said tumor cell and (i) a VH-domain which exclusively binds a carbohydrate epitope of said glycoprotein of said tumor cell, i.e. which does not contribute to or interfere with binding peptide epitope of said glycoprotein of said tumor cell; and wherein the peptide epitope of the glycoprotein of the tumor cell is associated with the carbohydrate epitope of said glycoprotein of said tumor cells, and wherein the antibody is specific for the combination of said carbohydrate epitope and said peptide epitope on said glycoprotein of said cancer cell.
  • the antibody of the present invention is specific for the combination of the carbohydrate epitope, which is a short truncated O-glycan, and the peptide epitope of a glycopeptide of a tumor cell, such as on the surface of the tumor cell.
  • the term "specific” in this regard refers to the specific antibody being highly selective for a particular glycoprotein, exhibiting strong binding and recognition for that glycoprotein, while displaying minimal or no binding to other types of glycoproteins. Specificity may be expressed by determining the binding affinity of the antibody to the glycoprotein, by using biophysical techniques as recognized by a person skilled in the art.
  • the antibodies of the present invention recognize both the carbohydrate moiety and the peptide sequence of the glycoprotein making them very specific.
  • the peptide epitope recognized by the VL-domain is part of a specific glycoprotein on the surface of a specific type of cancerous cells.
  • combotope antibodies of the invention is selected to be specific for the combined glycoprotein epitope, i.e. the carbohydrate epitope in combination with the peptide epitope (the combotope), and not for each of the epitopes individually.
  • the present invention provides an antibody which binds one or more tumor cells, wherein the one or more tumor cells comprises a carbohydrate epitope, preferably a short truncated O-glycan carbohydrate epitope, associated with a peptide epitope, wherein the antibody is specific against both the carbohydrate epitope and the peptide epitope.
  • the specific antibody does not bind specifically to combinations other than the specific combotope, and does not bind (or binds less effectively) the carbohydrate epitope or peptide epitope separately.
  • the present invention provides an antibody which binds a tumor cell, wherein the tumor cell comprises a carbohydrate epitope, preferably a short truncated O-glycan, associated with a peptide epitope, wherein the antibody is specific for the combination of said carbohydrate epitope and said peptide epitope by virtue of the antibody comprising
  • carbohydrate epitope is part of a glycoprotein and the peptide epitopes is part of said same glycoprotein.
  • the carbohydrate and peptide epitopes of the tumor are preferably surface displayed in order to facilitate recognition by the antibody.
  • the carbohydrate epitope to which the VH-domain binds is selected from mono-Tn, bis-Tn, mono-STn, bis-STn, and/or a combination of monoTn and monoSTn.
  • the carbohydrate epitope may comprise or consist of one or two adjacent short truncated O-glycans, i.e. Tn, TrnTn, STn, STrnSTn.
  • the present invention provides an antibody which binds a tumor cell, wherein the antibody comprises
  • a VL-domain which binds a peptide epitope of said plycoprotein of said tumor cells, wherein the peptide epitope of the glycoprotein of the tumor cell is associated with the carbohydrate epitope of said glycoprotein of the tumor cell, and wherein the antibody is specific for the combination of said carbohydrate epitope and said peptide epitope of said glycoprotein of said cancer cell.
  • the target peptide epitope of the cell is associated with a carbohydrate epitope.
  • the term "associated with” preferably refers to the peptide epitope and carbohydrate epitope being a continuous epitope, i.e. where the carbohydrate is directly attached to the peptide by virtue of being chemically linked, such as covalently linked.
  • the peptide epitope and carbohydrate epitope may be a discontinuous epitope, where the "associated with” still refers to the peptide epitope and carbohydrate epitope being in close proximity of one another, but by virtue of the structural configuration of the molecule facilitating this.
  • a discontinued epitope is where the amino acid hosting the attached carbohydrate epitope is not part of the peptide epitope.
  • the present invention provides a tumor cell binding antibody, comprising
  • a VH-domain which binds a carbohydrate epitope of a glycoprotein of the tumor cell preferably the carbohydrate epitope is one or two short truncated O- glycan(s)
  • the present invention provides a tumor cell binding antibody, comprising
  • a VH-domain which binds a Tn- and/or STn-carbohydrate epitope of a glycoprotein of the tumor cell preferably the carbohydrate epitope is one or two short truncated O-glycan(s)
  • the carbohydrate epitope is one or two short truncated O-glycan(s)
  • the carbohydrate epitope is composed of one or two short truncated O-glycan(s), where the a short truncated O-glycans are selected from Tn and STn.
  • a Tn or a STn epitope may be formed by one or two Tn or STn moieties.
  • the Tn or STn epitope which interacts with the VH domain is only one Tn or one STn moiety, respectively.
  • the Tn or STn epitope which interacts with the VH domain is only one Tn or one STn moiety, and said Tn or STn epitope is covalently attached to an amino acid residue, wherein said amino acid residue is part of the peptide epitope which interacts with the VL-domain.
  • the Tn or STn epitope is formed by two Tn or STn moieties, respectively, which both interacts with the VH domain.
  • the Tn or STn epitope which interacts with the VH domain is two Tn or one STn moieties, and said two Tn or STn moieties are covalently attached to two separate amino acid residue, wherein said amino acid residues are part of the peptide epitope which interacts with the VL-domain.
  • said two separate amino acid residues are adjacent to each other; in another embodiment said two separate amino acid residies are spaced apart by 1, 2, 3 or 4 other amino acid residues.
  • the VL-domain of the combotope antibody binds a peptide epitope of a glycoprotein of a tumor cell.
  • the peptide epitope is 2, 3, 4, 5, 6, 7, 8, 9, 10, 11, or 12 amino acid residues, preferably 2, 3, or 4 amino acid residues, most preferably 4 amino acid residues.
  • the peptide epitope which interacts with the VL domain is between 2-4, between 4-6, between 6-8, between 8-10, between 10-12 amino acid residues, such as between 2-12, between 2-10, between 2- 8, between preferably 2-6 amino acid residues, most preferably 2-4 amino acid residues.
  • the antibodies of the present invention are different from the prior art antibodies 5E5, 5F7, and 2D9.
  • 5F7 is disclosed in US11161911B2.
  • 5E5 is disclosed in W02008/040362 and US2021060070A1, and further in the scientific literature Macias-Leon et al 2020; Tarp et al 2007; and Blixt et al 2010.
  • 2D9 is disclosed in the scientific litterature Sorensen et al 2006; Tarp et al 2007; and Blixt et al 2010.
  • the antibodies of the present invention are different from the prior art antibodies 5E5, 5F7, and 2D9 - hence, the antibodies of the present invention do not comprise VL and VH domain combinations as disclosed here:
  • amino acid sequence of the antibodies of the present invention do not comprise an amino acid sequence combination selected from SEQ ID NO. 3 + 4, SEQ ID NO. 5 + 6, and SEQ ID NO. 7 + 8.
  • the antibody of the present invention is a humanized antibody, such as prepared by the method of Clavero-Alvarez et al 2018. "Humanized” forms of nonhuman antibodies can be chimeric antibodies that contain minimal sequence derived from the non-human antibody.
  • a humanized antibody is generally a human antibody (recipient antibody) in which selected residues in the non-human antibody (donor antibody) have been replaced.
  • the donor antibody can be any suitable non-human antibody, such as a mouse, rat, rabbit, chicken, or non-human primate antibody having a desired specificity, affinity, or biological effect.
  • the donor antibody is preferably identified by screening an antibody library of the present invention.
  • selected framework region residues of the recipient antibody are replaced by the corresponding framework region residues from the donor antibody.
  • Humanized antibodies may also comprise residues that are not found in either the recipient antibody or the donor antibody. In some instances, these modifications are made to further refine antibody performance. The skilled person would be familiar of methods to transform combotope antibodies of the present invention into humanized combotope antibodies.
  • the present invention provides nucleic acid sequences encoding the antibodies according to the present invention as disclosed herein.
  • the VH domain of the combotope antibody binds the carbohydrate epitope, preferably a mono-Tn, bis-Tn, mono-STn or bis-STN carbohydrate epitope.
  • the VH domain of the combotope antibody binds a Tn- carbohydrate epitope, such a mono-Tn or bis-Tn.
  • the inventors of the present invention surprisingly made the following discovery:
  • the VH-domain of antibody G2D11 (SEQ ID NO. 1) was by structural characterization in the presence of the bis-Tn-MUCl peptide APGS*T*AP (where * denotes a GalNAc moiety) found to recognize the two GalNAc moieties, but did not recognize the peptide sequence on which they were attached (see examples 1).
  • SEQ ID NO. 1 VH-domain G2D11: QVQMQQSDAELVKPGASVKISCKASGYIFADHAIHWVKRKPEQGLEWIGYISPGNDDIKYNEKF KGKATLTADKSSSTAYMQLNSLTSEDSAVYFCKRSLPGTFDYWGQGTTLTVSS
  • the novel combotope antibodies disclosed herein are partially based on these structural observations - i.e. that the VH-chain of G2D11 provides recognition support for the glycoside part of the antigen, without being affected/influenced by the peptide part of the glycoprotein and did not bind to this peptide.
  • the further development specifies a combotope antibody, where the VL-domain provides specific binding of the peptide epitope of the combotope and the VH-chain provides recognition support for the carbohydrate part of the glycoprotein.
  • several anti-Tn antibodies are known in the art, and many of these share the same germline sequences as the G2D11 (at least the essential sequences).
  • the prior art anti-Tn antibodies may comprise same germline sequences as the G2D11 responsible for the Tn-binding VH domain, but with no or an unknown binding contribution to the peptide/protein carrier - hence, they are unspecific Tn-binding antibodies, and therefore not combotope antibodies according to the present invention.
  • combotopes of the present invention are screened for by use of the antibody library of the present invention, as further disclosed herein, to obtain antibodies specific for a specific glycoproteins of interest.
  • the VH-domain of the combotope antibody of the present invention is a G2D11-Iike VH-domain.
  • the amino acids of the VH-domain of the combotope antibody of the present invention resemble the G2D11-Iike VH-domain in their structural conformation.
  • G2D11 tolerates binding of any combination Tn-Thr/Tn-Ser, Tn-Ser/Tn-Thr, Tn-Ser/Tn- Ser, Tn-Thr/Tn-Thr.
  • amino acid residues H32, A33, H35, Y50, and S99 are key residues for the binding of the VH domain to one of the GalNAc moiety of bis-Tn-MUCl on the peptide
  • amino acid residues S52, N55, and D57 are key residues for the binding of the VH domain to the other GalNAc moiety of bis-Tn-MUCl on the peptide.
  • the VH-domain of the combotope antibody of the present invention comprises an amino acid sequence having at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 1, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 1; and comprises the amino acid residues H32, A33, H35, Y50, and S99, with respect to SEQ ID NO. 1.
  • Said combotope antibody will provide recognition support for mono-Tn epitopes.
  • combotope antibodes for targeting tumor cells comprise a VH and a VL domain; wherein the VH domain of the antibodyis a Tn-binding domain and the amino acid sequence of the VH-domain has at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 1, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 1, wherein the amino acid sequence comprises the amino acid residues H32, A33, H35, Y50, and S99, with respect to SEQ ID NO.
  • VL domain of the antibody binds a peptide backbone epitope associated with the Tn-carbohydrate on the tumor cell.
  • the sequence of the VH domain is preferably such that there is no binding to the peptide epitope.
  • the VH-domain of the combotope antibody of the present invention comprises an amino acid sequence having at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 1, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 1; and comprises the amino acid residues S52, N55, and D57, with respect to SEQ ID NO. 1.
  • Said combotope antibody will provide recognition support for mono-Tn epitopes.
  • combotope antibodes for targeting tumor cells comprising a VH and a VL domain; wherein the VH domain of the antibody is a Tn-binding domain and the amino acid sequence of the VH-domain has at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 1, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 1, wherein the amino acid sequence comprises the amino acid residues S52, N55, and D57, with respect to SEQ ID NO.
  • VL domain of the antibody binds a peptide backbone epitope associated with the Tn- carbohydrate on the tumor cell.
  • the sequence of the VH domain is preferably such that there is no binding to the peptide epitope.
  • amino acid residues H32, A33, and H35 in CDR1 should preferably be conserved for the VH domain of the present invention; further, with reference to SEQ ID NO. 1, amino acid residues Y50, S52, N55, and D57 in the CDR2 should preferably be conserved for the VH domain of the present invention; and further, with reference to SEQ ID NO. 1, amino acid residue S99 in the CDR3 should preferably be conserved for the VH domain of the present invention.
  • the VH-domain of the combotope antibody comprises an amino acid sequence having at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 1, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 1; and comprises the amino acid residues H32, A33, H35, Y50, S52, N55, D57, and S99, with respect to SEQ ID NO. 1.
  • Said combotope antibody will also provide recognition support for the bis-Tn epitope.
  • the amino acid sequence of the VH-domain of the combotope antibody has at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 1, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 1; and comprises the amino acid residues H32, A33, H35, Y50, S52, N55, D57, and S99, with respect to SEQ ID NO. 1.
  • the VH domain of the combotope antibody comprises an amino acid sequence having at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 1, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 1; and in pairwise alignment with SEQ ID NO.
  • the amino acid sequence of the VH-domain comprises amino acid residues histidine (H), alanine (A), histidine (H), tyrosine (Y), serine (S), asparagine (N), aspartic acid (D), and serine (S) at positions corresponding to amino acid position H32, A33, H35, Y50, S52, N55, D57, and S99 of SEQ ID NO. 1, respectively.
  • the pairwise sequence alignment is performed using scoring matrix: blosum62, gap opening penalty: 10, and gap extension penalty 0.2.
  • the amino acid sequence of the VH-domain of the combotope antibody comprises amino acid residues histidine (H), alanine (A), histidine (H), tyrosine (Y), serine (S), asparagine (N), aspartic acid (D), and serine (S), at positions corresponding to amino acid position H32, A33, H35, Y50, S52, N55, D57, and S99, of SEQ ID NO. 1, respectively; and the amino acid sequence of the VH domain has at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 1, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 1.
  • antibodes for targeting tumor cells comprising a VH and a VL domain; wherein the VH domain of the antibody is a Tn-binding domain and the amino acid sequence of the VH-domain has at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 1, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 1, wherein the amino acid sequence comprises the amino acid residues H32, A33, H35, Y50, S52, N55, D57, and S99, with respect to SEQ ID NO.
  • the VL domain of the antibody binds a peptide backbone epitope associated with the Tn-carbohydrate on the tumor cell.
  • the sequence of the VH domain is preferably such that there is no binding to the peptide epitope.
  • Combotope antibodies of the present invention as described herein, such as the Tn- combotopes disclosed herein, comprise improved binding affinity to a specific antigen epitope termed a combotope comprised of two different epitopes on a glycoprotein.
  • the antibody comprises a binding affinity (e.g. kD) of between 100 nM to IpM, such as less than 100 nM, less than 10 nM, less than 1 nM, less than 100 pM, or even less than 10 pM.
  • the combotope antibodies of the present invention are used to treat cancer.
  • the cancer is lung, head and neck squamous cell, colorectal, melanoma, liver, classical Hodgkin lymphoma, kidney, gastric, cervical, merkel cell, B-cell lymphoma, or bladder cancer.
  • the cancer is a solid tumor.
  • the antibody comprises
  • VH-domain having at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 1, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 1, and comprising the amino acid residues H32, A33, H35, Y50, S52, N55, D57, and S99, with respect to SEQ ID NO. 1, and
  • VL domain having an amino acid sequence selected from of any one of SEQ ID NOs. 9-21.
  • the present invention provides an antibody comprising a VH domain having the amino acid sequence of SEQ ID NO. 1 and a VL domain having an amino acid sequence seleted from any one of SEQ ID NO. 9-21.
  • the present invention provides antibodies, wherein the antibody comprises (i) a VH-domain having at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 1, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 1, and comprising the amino acid residues H32, A33, H35, Y50, S52, N55, D57, and S99, with respect to SEQ ID NO.
  • the antibody is a monoclonal antibody, a polyclonal antibody, a bi
  • an antibody comprising (i) a VH domain having an amino acid sequence of SEQ ID NO. 1 and (ii) a VL domain having an amino acid sequence of any one of SEQ ID NOs. 9-21 is used to treat cancer.
  • the cancer is lung, head and neck squamous cell, colorectal, melanoma, liver, classical Hodgkin lymphoma, kidney, gastric, cervical, merkel cell, B-cell lymphoma, or bladder cancer.
  • the cancer is a solid tumor.
  • an antibody comprising (i) a VH domain having an amino acid sequence of SEQ ID NO. 1 and (ii) a VL domain having an amino acid sequence of any one of SEQ ID NOs. 9-21 is used to diagnose a cancerous state.
  • the antibody may be used for in vivo diagnosis or for identifying a cancerous state ex vivo in a tissue or cell sample taken from a patient.
  • Binding of the combotope antibody to a cancer target for diagnosis can be monitored or identified by way of known methods, such as for example labelling of the combotope antibodies and/or by use of labelled antibodies for binding of the combotope antibodies.
  • SEQ ID NO. 15 (VL-domain F4-TnCD43) : DYKDIQMTQSPASLSASVGETVTITCRASENIYSYLAWYQQKQGKSPQLLVYNAKTLAEGVPSRF SGSGSGTQYSLKINSLQPEDFGSYYCQHFWSTPYTFGGGTKLEMKR
  • the selected antibodies of the present invention are a humanized antibody, such as mententioned above.
  • the present invention provides a humanized antibody comprising a VH domain having the amino acid sequence of SEQ ID NO. 159 and a VL domain having an amino acid sequence of SEQ ID NO. 160.
  • a humanized antibody comprising a VH domain having the amino acid sequence of SEQ ID NO. 159 and a VL domain having an amino acid sequence of SEQ ID NO. 160 is used to treat cancer, e.g. a solid tumor cancer.
  • a humanized antibody comprising a VH domain having the amino acid sequence of SEQ ID NO. 159 and a VL domain having an amino acid sequence of SEQ ID NO. 160 is used to diagnose a cancerous state, e.g. by in vivo diagnosis or for identifying a cancerous state in a tissue or cell sample taken from a patient.
  • the present invention provides a humanized antibody comprising a VH domain having the amino acid sequence of SEQ ID NO. 159 and a VL domain having an amino acid sequence of SEQ ID NO. 161.
  • a humanized antibody comprising a VH domain having the amino acid sequence of SEQ ID NO. 159 and a VL domain having an amino acid sequence of SEQ ID NO. 161 is used to treat cancer, e.g. a solid tumor cancer.
  • a humanized antibody comprising a VH domain having the amino acid sequence of SEQ ID NO. 159 and a VL domain having an amino acid sequence of SEQ ID NO. 161 is used to diagnose a cancerous state, e.g. by in vivo diagnosis or for identifying a cancerous state in a tissue or cell sample taken from a patient.
  • the present invention provides a humanized antibody comprising a VH domain having the amino acid sequence of SEQ ID NO. 162 and a VL domain having an amino acid sequence of SEQ ID NO. 163.
  • a humanized antibody comprising a VH domain having the amino acid sequence of SEQ ID NO. 162 and a VL domain having an amino acid sequence of SEQ ID NO. 163 is used to treat cancer, e.g. a solid tumor cancer.
  • a humanized antibody comprising a VH domain having the amino acid sequence of SEQ ID NO. 162 and a VL domain having an amino acid sequence of SEQ ID NO. 163 is used to diagnose a cancerous state, e.g. by in vivo diagnosis or for identifying a cancerous state in a tissue or cell sample taken from a patient.
  • the present invention provides a humanized antibody comprising a VH domain having the amino acid sequence of SEQ ID NO. 164 and a VL domain having an amino acid sequence of SEQ ID NO. 165.
  • a humanized antibody comprising a VH domain having the amino acid sequence of SEQ ID NO. 164 and a VL domain having an amino acid sequence of SEQ ID NO. 165 is used to treat cancer, e.g. a solid tumor cancer.
  • a humanized antibody comprising a VH domain having the amino acid sequence of SEQ ID NO. 164 and a VL domain having an amino acid sequence of SEQ ID NO. 165 is used to diagnose a cancerous state, e.g. by in vivo diagnosis or for identifying a cancerous state in a tissue or cell sample taken from a patient.
  • the present invention provides a humanized antibody comprising a VH domain having the amino acid sequence of SEQ ID NO. 166 and a VL domain having an amino acid sequence of SEQ ID NO. 167.
  • a humanized antibody comprising a VH domain having the amino acid sequence of SEQ ID NO. 166 and a VL domain having an amino acid sequence of SEQ ID NO. 167 is used to treat cancer, e.g. a solid tumor cancer.
  • a humanized antibody comprising a VH domain having the amino acid sequence of SEQ ID NO. 166 and a VL domain having an amino acid sequence of SEQ ID NO. 167 is used to diagnose a cancerous state, e.g. by in vivo diagnosis or for identifying a cancerous state in a tissue or cell sample taken from a patient. I.ii STn-combotopes
  • the VH domain of the combotope antibody binds the carbohydrate epitope, preferably a mono-Tn, bis-Tn, mono-STn or bis-STn carbohydrate epitope.
  • the VH domain of the combotope antibody binds a STn- carbohydrate epitope, such a mono-STn or bis-STn.
  • the VH-domain of antibody G2D11 was by structural comparison to the VH domain of 3F1 (SEQ ID NO. 25) modified into a STn-binding VH-domain (SEQ ID NO. 28) (see example 5).
  • the STn binding VH domain (SEQ ID NO. 28) has the following amino acid residue changes: I28T, A30T, P101L, de/G102, T103A and F104L.
  • SEQ ID NO. 28 VH domain G2D11 mutant M2: LAL-TFT
  • novel STn-combotope antibodies disclosed herein are based on these structural modification and further development, as disclosed herein, and provides specific combotope recognition, wherein the VH-chain provides recognition support for the carbohydrate part STn of the combotope on the glycoprotein antigen, while the VL- domain provides recognition support for the peptide part of the combotope on the glycoprotein antigen.
  • the VH-domain of the combotope antibody of the present invention is a SEQ ID NO. 28-like VH-domain.
  • the amino acids of the VH- domain of the combotope of the present invention resemble the SEQ ID NO. 28-like VH- domain in their structural conformation.
  • amino acid residues T28, T30, L101, A102, and L103 with respect to SEQ ID NO. 28 are key amino acids for the STn-specificity. They are need for accommodation of the sialyl - i.e. for making extra space for sialyl for STn binding.
  • the VH-domain of the combotope antibody comprises an amino acid sequence having at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 28, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 28; and comprises the amino acid residues T28, T30, H32, A33, H35, Y50, S99, L101, A102 and L1O3, with respect to SEQ ID NO. 28.
  • Said combotope antibody will provide recognition support for mono-STn epitopes.
  • antibodies for targeting tumor cells are provided, said antibody comprsing a VH and VL domain; wherein the VH domain of the antibody is a STn-binding domain and the amino acid sequence of the VH-domain has at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 28, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 28; wherein the amino acid sequence comprises the amino acid residues T28, T30, H32, A33, H35, Y50, S99, L101, A102 and L103, with respect to SEQ ID NO. 28; and wherein the VL domain of the antibody binds a peptide backbone associated with the STn-carbohydrate on the tumor cell.
  • the sequence of the VH domain is preferably such that there is no binding to the peptide epitope.
  • the VH-domain of the combotope antibody comprises an amino acid sequence having at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 28, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 28; and comprises the amino acid residues T28, T30, S52, N55, D57, L101, A102 and L103, with respect to SEQ ID NO. 28.
  • Said combotope antibody will provide recognition support for the mono- STn epitope.
  • antibodes for targeting tumor cells comprising a VH and VL domain; wherein the VH domain of the antibody is a STn-binding domain and the amino acid sequence of the VH-domain has at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 28, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 28; wherein the amino acid sequence comprises the amino acid residues T28, T30, S52, N55, D57, L101, A102 and L103, with respect to SEQ ID NO. 28; and wherein the VL domain of the antibody binds a peptide backbone associated with the STn-carbohydrate on the tumor cell.
  • the sequence of the VH domain is preferably such that there is no binding to the peptide epitope.
  • the VH-domain of the combotope antibody comprises an amino acid sequence having at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 28, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 28; and comprises the amino acid residues T28, T30, H32, A33, H35, Y50, S52, N55, D57, S99, L101, A102 and L103, with respect to SEQ ID NO. 28.
  • Said combotope antibody will also provide recognition support for bis-STn epitopes.
  • the amino acid sequence of the VH-domain of the combotope antibody has at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 28, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 28; and comprises the amino acid residues T28, T30, H32, A33, H35, Y50, S52, N55, D57, S99, L101, A102 and L103, with respect to SEQ ID NO. 28.
  • the VH domain of the combotope antibody comprises an amino acid sequence having at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 28, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 28; and in pairwise alignment with SEQ ID NO.
  • the amino acid sequence of the VH-domain comprises amino acid residues threonine (T), threonine (T), histidine (H), alanine (A), histidine (H), tyrosine (Y), serine (S), asparagine (N), aspartic acid (D), serine (S), leucine (L), alanine (A), and leucine (L) at positions corresponding to amino acid position T28, T30, H32, A33, H35, Y50, S52, N55, D57, S99, L101, A102 and L103 of SEQ ID NO. 28, respectively.
  • the pairwise sequence alignment is performed using scoring matrix: blosum62, gap opening penalty: 10, and gap extension penalty 0.2.
  • the amino acid sequence of the VH-domain of the combotope antibody comprises amino acid residues threonine (T), threonine (T), histidine (H), alanine (A), histidine (H), tyrosine (Y), serine (S), asparagine (N), aspartic acid (D), serine (S), leucine (L), alanine (A), and leucine (L), at positions corresponding to amino acid position T28, T30, H32, A33, H35, Y50, S52, N55, D57, S99, L101, A102 and L103, of SEQ ID NO.
  • amino acid sequence of the VH domain has at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 28, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 28.
  • antibodes for targeting tumor cells comprising a VH and VL domain; wherein the VH domain of the antibody is a STn-binding domain and the amino acid sequence of the VH-domain has at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 28, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO.
  • amino acid sequence comprises the amino acid residues T28, T30, H32, A33, H35, Y50, S52, N55, D57, S99, L101, A102 and L103, with respect to SEQ ID NO. 28; and wherein the VL domain of the antibody binds a peptide backbone associated with the STn-carbohydrate on the tumor cell.
  • the sequence of the VH domain is preferably such that there is no binding to the peptide epitope.
  • Combotope antibodies of the present invention as described herein, such as the STn- combotopes disclosed herein, comprise improved binding affinity to a specific antigen epitope, the combotope (compared to other epitopes on healthy or cancer cells).
  • the antibody comprises a binding affinity (e.g., kD) of between 100 nM to IpM, such as less than 100 nM, less than 10 nM, less than 1 nM, less than 100 pM, or even less than 10 pM..
  • the antibodies of the present invention are used to treat cancer.
  • the cancer is lung, head and neck squamous cell, colorectal, melanoma, liver, classical Hodgkin lymphoma, kidney, gastric, cervical, merkel cell, B- cell lymphoma, or bladder cancer.
  • the cancer is a solid tumor.
  • an antibody comprising (i) a VH domain having an amino acid sequence of SEQ ID NO. 28 and (ii) a VL domain having an amino acid sequence of any one of SEQ ID NOs. 9-21 is used to diagnose a cancerous state.
  • the antibody may be used for in vivo diagnosis or for identifying a cancerous state ex vivo in a tissue or cell sample taken from a patient.
  • Binding of the combotope antibody to a cancer target for diagnosis can be monitored or identified by way of known methods, such as for example labelling of the combotope antibodies and/or by use of labelled antibodies for binding of the combotope antibodies.
  • the present invention discloses specific antibodies.
  • the combotope antibody of the present invention comprises
  • VH-domain having at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 28, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 28, and comprising the amino acid residues T28, T30, H32, A33, H35, Y50, S52, N55, D57, S99, L101, A102 and L103, with respect to SEQ ID NO. 28, and
  • a VL domain having an amino acid sequence selected from of any one of SEQ ID NOs. 22-24 (ii) a VL domain having an amino acid sequence selected from of any one of SEQ ID NOs. 22-24.
  • the present invention provides an antibody comprising a VH domain having the amino acid sequence of SEQ ID NO. 28 and a VL domain having an amino acid sequence seleted from any one of SEQ ID NO. 22-24.
  • the present invention provides antibodies, wherein the antibody comprises (i) a VH-domain having at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 28, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 28, and comprising the amino acid residues T28, T30, H32, A33, H35, Y50, S52, N55, D57, S99, L101, A102 and L103, with respect to SEQ ID NO. 28, and (ii) a VL domain having an amino acid sequence selected from of any one of SEQ ID NOs.
  • the antibody is a monoclonal antibody, a polyclonal antibody, a bi-specific antibody, a multispecific antibody, a grafted antibody, a human antibody, a humanized antibody, a synthetic antibody, a chimeric antibody, a camelized antibody, a single-chain Fvs (scFv), a single chain antibody, a Fab fragment, a F(ab')2 fragment, a Fd fragment, a Fv fragment, a single-domain antibody, an isolated complementarity determining region (CDR), a diabody, a fragment comprised of only a single monomeric variable domain, disulfide-linked Fvs (sdFv), an intrabody, an anti-idiotypic (anti-Id) antibody, or ab antigen-binding fragments thereof.
  • scFv single chain antibody
  • Fab fragment a F(ab')2 fragment
  • Fd fragment fragment
  • a single-domain antibody an isolated complementarity determining region (CDR)
  • an antibody comprising (i) a VH domain having an amino acid sequence of SEQ ID NO. 28 and (ii) a VL domain having an amino acid sequence of any one of SEQ ID NOs. 22-24 is used to treat cancer.
  • the cancer is lung, head and neck squamous cell, colorectal, melanoma, liver, classical Hodgkin lymphoma, kidney, gastric, cervical, merkel cell, B-cell lymphoma, or bladder cancer.
  • the cancer is a solid tumor.
  • the selected antibody of the present invention is a humanized antibody, such as mententioned above.
  • V.iii Tn-combotopes or STn-combotopesIt may be the case that it is not known whether the combotope comprises Tn or STn as the short truncated O-glycan on the surface of a glycoprotein or that a particular glycoprotein comprisees a mixture or these glycans. In such a case, it may be beneficial to use an antibody that binds either a Tn or a STn epitope in the combotope.
  • the combotope antibody comprises both a VH domain for the Tn epitope and a VH domain for the STn epitope.
  • Such antibodies comprise a VH(Tn)-VLl arm and a VH(STn)- VL2 arm, where the VL1 and VL2 domains may be the same or different and for binding the same or different peptide epitope(s).
  • the VH(Tn)-VLl arm and the VH(STn)-VL2 arm may be comprised as the two Fab arms in a normal antibody, as F(ab)2, as a minibody, a diabody, a triabody or be two scFv linked in one molecule, e.g. scFv-Fc .
  • the VH(Tn) domain has an amino acid sequence of SEQ NO. 1 or a functional variant thereof as defined above and the VH(STn) has an amino acid sequence of SEQ NO. 28 or a functional variant thereof as defined above.
  • the present invention provides an antibody library for in-vitro identification of a specific antibody which binds glycoproteins, such as glycoproteins on cancer tumor cells.
  • the present invention provides an antibody library for in-vitro identification of a specific antibody which binds one or more tumor cells.
  • tumor cells often comprise Tn or STn glycosylation epitopes on specific glycoproteins on the surface of the cancer cells.
  • tumor cells often comprise short truncated O-glycans on the surface, exposing for example Tn and/or STn epitopes on specific glycoproteins on the surface of the cancer cells.
  • the antibody library of the present invention facilitates identification of antibodies with improved specificity towards tumor cells by virtue of the antibodies being specific both towards the carbohydrate epitope (Tn and/or STn) as well as towards the peptide epitope in the protein backdone associated with the carbohydrate (epitope).
  • such improved antibodies comprise a VH domain which efficiently binds the carbohydrate epitope of the glycoprotein, and a VL domain which efficiently binds the peptide epitope of the glycoprotein associated with the carbohydrate epitope, and the antibody is thereby specific for the combination of said carbohydrate epitope and said peptide epitope of said glycoprotein, the combined epitopes being termed a "combotope”.
  • VH-domain of antibody G2D11 was by structural characterization found to particularly recognize the carbohydrate epitope of the glycoprotein epitope bis-Tn on MUC1, while it did not recognize the peptide sequence of said glycoprotein epitope (see examples 1).
  • each antibody of the library comprises a specific preselected VH chain which provides recognition support for the carbohydrate part of the combotope antigen, while the VL-domain is variable, including one or more VL domains being specific for a specific peptide sequence of a particular glycoprotein, creating a library which can be screened for specific combotope antibodies, of which the VL-domain will provide recognition support to the peptide epitope within the combotope of said glycoprotein.
  • the antibody library of the present invention can be used to identify specific glycoprotein combotope antibodies, where the VL-domain is specific for the peptide epitope, i.e. glycoprotein, of choice.
  • the invention provides an antibody library, wherein each of the antibodies in the library comprises two antibody domains: (i) a first antibody domain which binds the carbohydrate epitope of a glycoprotein on a cancer cell, i.e. Tn, bis. Tn, STn or bis.
  • a second antibody domain selected from a repertoire of antibody domains, wherein the repertoire of antibody domains comprises one or more domains which binds a peptide epitope of said glycoprotein; wherein said library is for in-vitro identification of a specifc antibody from said library for targeting cancer cells, and wherein the specific antibody is specific for the combination of said carbohydrate epitope and said peptide epitope of said glycoprotein - i.e. the specifc antibody is a combotope antibody.
  • the invention provides an antibody library, wherein each of the antibodies in the library comprises two antibody domains: (i) a first antibody domain which binds the carbohydrate epitope of a glycoprotein of the tumor cell, and (ii) a second antibody domain selected from a repertoire of antibody domains, wherein the repertoire of antibody domains comprises one or more domains which binds a peptide epitope of said glycoprotein of said tumor cell; wherein said library is for in-vitro identification of a specific antibody from said library for targeting tumor cells, and wherein the specific antibody is specific for the combination of said carbohydrate epitope and said peptide epitope of said glycoprotein - i.e. the specifc antibody is a combotope antibody.
  • the first antibody domain is a VH domain and the second antibody domain is a VL domain. In another embodiment, both the first and second antibody are VH domains, but different from one another.
  • the present invention provides an antibody library, wherein each of the antibodies in the antibody library comprises
  • VL-domain selected from a repertoire of VL-domains, wherein the repertoire of VL-domains comprises one or more VL-domains which binds a peptide epitope on said glycoprotein of the tumor cell, for in-vitro identification of a specific antibody from said library which binds a tumor cell, wherein the specific antibody is specific for the combination of said carbohydrate epitope and said peptide epitope of said glycoprotein, the combined epitope being termed a "combotope".
  • the present invention provides an antibody library for in-vitro identification of a specific antibody which binds a tumor cell
  • each of the antibodies in the antibody library comprises (ii) a VL-domain selected from a repertoire of VL-domains, wherein the repertoire of VL-domains comprises one or more VL-domains which binds a peptide epitope on a glycoprotein of a tumor cell, and (i) a VH-domain which binds a carbohydrate epitope on said glycoprotein of said tumor cell and which does not contribute or interfere with binding said peptide epitope of said glycoprotein of said turner cell, wherein the specific antibody is specific for the combination of said carbohydrate epitope and said peptide epitope of said glycoprotein, the combined epitope being termed a "combotope" i.e. the specifc antibody is a combotope antibody.
  • the VH chain of each antibody encoded by the library thereby pre-selects for a desired glycoform specificity, while the VL chain is selected from a repertoire of LV domains and will determine the peptide backbone specificity and thereby the glycoprotein specificity.
  • the VH chain of each antibody encoded by the library pre-selects for a desired glycoform specificity, without interfering with the peptide backbone specificity, while the VL chain is selected from a repertoire of LV domains and will determine the peptide backbone specificity and thereby the glycoprotein specificity.
  • Example 1 The structural studies disclosed herein (Example 1) provide clear evidence that the VH domain of G2D11 has no interaction with the peptide/protein carrier, and that peptide/protein interaction of the antibodies entirely comes from contribution via the VL-chain, as exemplified with antibodies obtain from the library of the present invention (e.g. antibodies Tn-MUCl and Tn-CD43; see Examples 3 and 4).
  • the peptide epitope and carbohydrate epitope are a common continuous epitope, i.e. where the carbohydrate is in close proximity to the peptide by virtue of being chemically linked, such as by a covalent bond.
  • the peptide epitope and carbohydrate epitope may be a common discontinuous epitope, where the "associated with” still refers to the peptide epitope and carbohydrate epitope being in close proximity of one another, but by virtue of the structural configuration of the molecule facilitating the common epitope.
  • the antibodies herein are selected from a monoclonal antibody, a polyclonal antibody, a bi-specific antibody, a multispecific antibody, a grafted antibody, a human antibody, a humanized antibody, a synthetic antibody, a chimeric antibody, a camelized antibody, a single-chain Fvs (scFv), a single chain antibody, a Fab fragment, a F(ab')2 fragment, a Fd fragment, a Fv fragment, a single-domain antibody, an isolated complementarity determining region (CDR), a diabody, a fragment comprised of only a single monomeric variable domain, disulfide-linked Fvs (sdFv), an intrabody, an antiidiotypic (anti-Id) antibody, and ab antigen-binding fragments thereof.
  • scFv single chain antibody
  • Fab fragment a F(ab')2 fragment
  • Fd fragment fragment
  • a single-domain antibody an isolated complementarity determining region (C
  • the antibodies encoded by the antibody library of the invention are scFv, wherein the VH-domain is linked to the VL-domain.
  • the libraries disclosed herein comprise scFv antibodies, comprising a VH domain and a VL domain, where both domains are present in a single polypeptide chain.
  • the Fv polypeptide further comprises a polypeptide linker between the VH and VL domains allowing the scFv to form the desired structure for antigen binding.
  • the linker is selected from (GGGGS) n where in is 1, 2, 3, 4, 5, or 6. Many other linkers could be used as an alternative to the exemplified (GGGGS) n -linker. The skilled person would know how to select such linkers.
  • the VH domain of each antibody in the antibody library specifically binds one or more carbohydrate epitope selected from Tn and/or STn. In one embodiment, the VH domain of each antibody encoded by the antibody library specifically binds one or more Tn moieties, such as a mono-Tn epitope or a bis-Tn epitope. In another embodiment, the VH domain of each antibody encoded by the antibody library specifically binds one or more STn-moieties, such as a mono-STn epitope or a bis-STn epitope.
  • the VH domain of some of the antibodies encoded by the antibody library specifically bind one or more Tn-moieties, while the VH domain of other antibodies encoded by the antibody library specifically bind one or more STn-moieties.
  • the antibody comprises two different VH domains, one for Tn or bis-Tn and the other for STn or bis-STn.
  • the VH domain is specified for each antibody in the antibody library to specifically bind a specific carbohydrate epitope being characteristic of cancer cells.
  • the VL domains of the antibodies in the antibody library vary from one antibody to the other, representing a repertoire of VL domains recognizing different peptides from the glycosylated protein in question, such that a repertoire of VL-domains is generated, to be screened with the intent of identifying combotope antibodies which have specificity towards glycoproteins on tumor cells by virtue of the VH domain binding a carbohydrate epitope of the glycoprotein on the tumor cell and the VL domain binding a peptide epitope on said glycoprotein on the tumor cell.
  • the repertoire of VL-domains is generated from a naive immune repertoire, an immunized immune repertoire in a suitable animal, or a synthetically produced repertoire.
  • the repertoire of VL-domains encoded by the antibody library is a naive immune repertoire from an animal, such as a mouse or human.
  • suitable animals are pigs, rats, dogs, horses, rabbits known to the skilled artisan.
  • the antibody library is phage display library.
  • Tn-template antibody libraries wherein the first domain of each antibody in the library is a Tn-binding domain, while the second domain is selected from a repertoire of antibody domains comprising one or more domain binding a peptide epitope on a glycoprotein of interest associated with the Tn epitope, as disclosed above.
  • Tn-template antibody libraries wherein the VH domain of each antibody in the library is a Tn-binding VH domain, while the VL domain is selected from a repertoire of VL domains comprising one or more VL-domains binding a peptide epitope on a glycoprotein of interest associated with the Tn epitope, as disclosed above.
  • the present invention provides mono-Tn-template antibody libraries, wherein the VH domain of each antibody in the library is a mono-Tn-binding VH domain, while the VL domain is selected from a repertoire of VL domains comprising one or more VL-domains binding a peptide epitope on a glycoprotein of interest associated with the Tn epitope (supra').
  • the present invention provides bis-Tn-template antibody libraries, wherein the VH domain of each antibody in the library is a bis-Tn-binding VH domain, while the VL domain is selected from a repertoire of VL domains comprising one or more VL-domains binding a peptide epitope on a glycoprotein of interest associated with the Tn epitope (supra).
  • the present invention provides antibody libraries which may be used to select for antibodies which binds mono-Tn and bis-Tn epitopes, wherein the VH domain of each antibody in the library is a mono- as well as bis-Tn-binding VH domain, while the VL domain is selected from a repertoire of VL domains comprising one or more VL-domains binding a peptide epitope on glycoprotein of interest associated with the Tn epitope (supra).
  • the VH-domain of each antibody in the Tn-template antibody library is a G2D11-Iike VH-domain (SEQ ID NO. 1).
  • the amino acids of the VH-domain of each antibody in the library resemble the G2D11-Iike VH-domain in their structural conformation.
  • the VH-domain of the Tn-template antibody library comprises an amino acid sequence having at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 1, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 1; and comprises amino acid residues H32, A33, H35, Y50, and S99, with respect to SEQ ID NO. 1.
  • an antibody library for selecting tumor-targeting antibodies, wherein each antibody of the antibody library comprises a VH and a VL domain; wherein the VH domain of each antibody is a Tn-binding VH domain and the VH-domain comprises an amino acid sequence having at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 1, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 1, and comprises amino acid residues H32, A33, H35, Y50, and S99, with respect to SEQ ID NO.
  • the VH domain is selected from a repertoire of VL domains, wherein the repertoire comprises at least one VL domain which binds a peptide backbone epitope associated with the Tn-carbohydrate on the tumor cell.
  • the sequence of the VH domain is preferably such that there is no binding to the peptide epitope.
  • the VH-domain of the Tn-template antibody library comprises an amino acid sequence having at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 1, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 1; and comprises amino acid residues S52, N55, and D57, with respect to SEQ ID NO. 1.
  • an antibody library for selecting tumor-targeting antibodies, wherein each antibody of the antibody library comprises a VH and a VL domain; wherein the VH domain of each antibody is a Tn-binding VH domain and the VH-domain comprises an amino acid sequence having at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 1, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 1, and comprises amino acid residues S52, N55, and D57, with respect to SEQ ID NO.
  • VL domain is selected from a repertoire of VL domains, wherein the repertoire comprises at least one VL domain which binds a peptide backbone epitope associated with the Tn-carbohydrate on the tumor cell.
  • the sequence of the VH domain is preferably such that there is no binding to the peptide epitope.
  • the VH-domain of the Tn-template antibody library comprises an amino acid sequence having at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 1, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 1; and comprises amino acid residues H32, A33, H35, Y50, S52, N55, D57, and S99, with respect to SEQ ID NO. 1.
  • an antibody library for selecting tumortargeting antibodies, wherein each antibody of the antibody library comprises a VH and a VL domain; wherein the VH domain of each antibody is a Tn-binding VH domain and the VH-domain comprises an amino acid sequence having at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 1, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 1, and comprises amino acid residues H32, A33, H35, Y50, S52, N55, D57, and S99, with respect to SEQ ID NO.
  • the VL domain is selected from a repertoire of VL domains, wherein the repertoire comprises at least one VL domain which binds a peptide backbone epitope associated with the Tn-carbohydrate on the tumor cell.
  • the sequence of the VH domain is preferably such that there is no binding to the peptide epitope.
  • the VH domain of the Tn-template antibody library comprises an amino acid sequence having at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 1, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 1; and in pairwise alignment with SEQ ID NO.
  • the amino acid sequence of the VH-domain comprises amino acid residues histidine (H), alanine (A), histidine (H), tyrosine (Y), serine (S), asparagine (N), aspartic acid (D), and serine (S) at positions corresponding to amino acid position H32, A33, H35, Y50, S52, N55, D57, and S99 of SEQ ID NO. 1, respectively.
  • the pairwise sequence alignment is performed using scoring matrix: blosum62, gap opening penalty: 10, and gap extension penalty 0.2.
  • an antibody library for selecting tumor-targeting antibodies, wherein each antibody of the antibody library comprises a VH and VL domain; wherein the VH domain of each antibody is a Tn-binding VH domain and the amino acid sequence of the VH-domain in pairwise alignment with SEQ ID NO. 1 comprises amino acid residues histidine (H), alanine (A), histidine (H), tyrosine (Y), serine (S), asparagine (N), aspartic acid (D), and serine (S) at positions corresponding to amino acid position H32, A33, H35, Y50, S52, N55, D57, and S99 of SEQ ID NO.
  • amino acid sequence of the VH domain has at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 1, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 1.
  • the first antibody domain is a mono- or bis-Tn-binding VH domain
  • said VH-domain comprises an amino acid sequence having at least 90% sequence homology to SEQ ID NO. 1 and comprising amino acid residues H32, A33, H35, Y50, and S99 and/or amino acid residues S52, N55, and D57, with respect to SEQ ID NO. 1, preferably comprising amino acid residues H32, A33, H35, Y50, and S99 and amino acid residues S52, N55, and D57, with respect to SEQ ID NO. 1.
  • the Tn-template library is useful for screening for antibodies which bind combotopes comprising a Tn epitope.
  • STn-template antibody libraries wherein the first domain of each antibody in the library is a STn-binding domain, while the second domain is selected from a repertoire of antibody domains comprising one or more domain binding a peptide epitope on a glycoprotein of interest associated with the STn epitope, as disclosed above.
  • STn-template antibody libraries wherein the VH domain of each antibody in the library is a STn-binding VH domain, while the VL domain is selected from a repertoire of VL domains comprising one or more VL-domains binding a peptide epitope on a glycoprotein of interest associated with the STn epitope, as describes above.
  • the present invention provides mono-STn-template antibody libraries, wherein the VH domain of each antibody in the library is a mono-STn-binding VH domain, while the VL domain is selected from a repertoire of VL domains comprising one or more VL-domains binding a peptide epitope on a glycoprotein of interest associated with the STn epitope (supra').
  • the present invention provides bis-STn-template antibody libraries, wherein the VH domain of each antibody in the library is a bis-STn-binding VH domain, while the VL domain is selected from a repertoire of VL domains comprising one or more VL-domains binding a peptide epitope on a glycoprotein of interest associated with the STn epitope (supra).
  • the present invention provides antibody libraries which may be used to select for antibodies which binds mono-STn and bis-STn epitopes, wherein the VH domain of each antibody in the library is a mono- as well as bis-STn-binding VH domain, while the VL domain is selected from a repertoire of VL domains comprising one or more VL-domains binding a peptide epitope on a glycoprotein of interest associated with the STn epitope (supra).
  • the VH-domain of antibody 3F1 efficienly binds STn.
  • the VH-domain of each antibody in the STn-template antibody library is a 3F1-Iike VH-domain (SEQ ID NO. 25).
  • the amino acids of the VH-domain of each antibody in the library resemble the 3F1-Iike VH-domain in their structural conformation.
  • a potential disadvantage of 3F1 is that is does not express well, however, G2D11 is very stable and easy to produce compared to 3F1. For example, G2D11 is very well expressed in Pichia pastoris, while 3F1 does not express so well. In addition, we wanted to learn the molecular basis of how to conver an anti-Tn to an anti-STn.
  • G2D11 and 3F1 amino acid residues were identified which might affect Tn vs STn specificity.
  • a modified G2D11 VH domain was prepared, to simulate the 3F1 VH domain, for obtaining STn specificity (see example 5).
  • This altered G2D11 VH domain is provided herein as SEQ ID NO. 28.
  • the altered STn binding VH domain has the following amino acid residue changes: I28T, A30T, P101L, de/G102, T103A and F104L.
  • the VH- domain of each antibody in the STn-template antibody library is a SEQ ID NO. 28-like VH-domain.
  • the amino acids of the VH-domain of each antibody in the library resemble SEQ ID NO. 28 in their structural conformation.
  • the VH-domain of the STn-template antibody library comprises an amino acid sequence having at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 28, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 28, and comprises amino acid residues T28, T30, H32, A33, H35, Y50, S99, L101, A102 and L103, with respect to SEQ ID NO. 28.
  • an antibody library for selecting tumor-targeting antibodies, wherein each antibody of the antibody library comprises a VH and a VL domain; wherein the VH domain of each antibody is a STn-binding VH domain and the VH-domain comprises an amino acid sequence having at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 28, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 28, and comprising amino acid residues T28, T30, H32, A33, H35, Y50, S99, L101, A102 and L103, with respect to SEQ ID NO. 28; and wherein the VL domain is selected from a repertoire of VL domains, wherein the repertoire comprises at least one VL domain which binds a peptide backbone epitope associated with the STn- carbohydrate on the tumor cell.
  • the VH-domain of the STn-template antibody library comprises an amino acid sequence having at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 28, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 28, and comprises amino acid residues T28, T30, S52, N55, D57, L101, A102 and L103, with respect to SEQ ID NO. 28.
  • an antibody library for selecting tumor-targeting antibodies, wherein each antibody of the antibody library comprises a VH and a VL domain; wherein the VH domain of each antibody is a STn-binding VH domain and the VH-domain comprises an amino acid sequence having at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 28, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 28, and comprising amino acid residues T28, T30, S52, N55, D57, L101, A102 and L103, with respect to SEQ ID NO. 28; and wherein the VL domain is selected from a repertoire of VL domains, wherein the repertoire comprises at least one VL domain which binds a peptide backbone epitope associated with the STn-carbohydrate on the tumor cell.
  • the VH-domain of the STn-template antibody library comprises an amino acid sequence having at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 28, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 28, and comprises amino acid residues T28, T30, H32, A33, H35, Y50, S52, N55, D57, S99, L101, A102 and L103, with respect to SEQ ID NO. 28.
  • an antibody library for selecting tumor-targeting antibodies, wherein each antibody of the antibody library comprises a VH and a VL domain; wherein the VH domain of each antibody is a STn-binding VH domain and the VH-domain comprises an amino acid sequence having at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 28, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO.
  • VL domain is selected from a repertoire of VL domains, wherein the repertoire comprises at least one VL domain which binds a peptide backbone epitope associated with the STn- carbohydrate on the tumor cell.
  • the VH domain of the STn-template antibody library comprises an amino acid sequence having at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 28, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 28, and in pairwise alignment with SEQ ID NO. 28, the amino acid sequence of the VH-domain comprises amino acid residues Thr, Thr, His, Ala, His, Tyr, Ser, Asn, Asp, Ser, Leu, Ala and Leu at positions corresponding to amino acid positions 28, 30, 32, 33, 35, 50, 52, 55, 57, 99, 101, 102 and 103 of SEQ ID NO. 28, respectively.
  • the pairwise sequence alignment is performed using scoring matrix: blosum62, gap opening penalty: 10, and gap extension penalty 0.2.
  • an antibody library for selecting tumor-targeting antibodies, wherein each antibody of the antibody library comprises a VH and a VL domain; wherein the VH domain of each antibody is a STn-binding VH domain and the amino acid sequence of the VH-domain in pairwise alignment with SEQ ID NO. 28 comprises amino acid residues Thr, Thr, His, Ala, His, Tyr, Ser, Asn, Asp, Ser, Leu, Ala and Leu at positions corresponding to amino acid positions 28, 30, 32, 33, 35, 50, 52, 55, 57, 99, 101, 102 and 103 of SEQ ID NO.
  • amino acid sequence of the VH domain has at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 28, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 28.
  • the first antibody domain is a mono- or bis-STn-binding VH domain
  • said VH-domain comprises an amino acid sequence having at least 90% sequence homology to SEQ ID NO. 28 and comprising amino acid residues (i), T28, T30, H32, A33, H35, Y50, S99, L101, A102 and L103, (ii), T28, T30, S52, N55, D57, L101, A102 and L103, or (iii) T28, T30, H32, A33, H35, Y50, S52, N55, D57, S99, L101, A102 and L103, with respect to SEQ ID NO. 28, preferably comprising amino acid residues T28, T30, H32, A33, H35, Y50, S52, N55, D57, S99, L101, A102 and L103, with respect to SEQ ID NO. 28.
  • the STn-template library is useful for screening for antibodies which bind combotopes comprising a STn epitope.
  • the library may further be useful for screening for antibodies which bind combotopes comprising a Tn epitope, or a combination of Tn and STn epitope.
  • the present invention provides a nucleic acid library encoding the antibody library disclosed herein.
  • the present invention provides a nucleic acid library encoding antibodies, wherein each of the nucleic acids in the library comprises
  • a second nucleic acid sequence selected from a repertoire of nucleic acids sequences comprising one or more nucleic acid sequences encoding an antibody domain that binds a peptide epitope of said glycoprotein of said tumor cell, for in-vitro identification of a specific antibody from said library which specifically binds a tumor cell, wherein the peptide epitope of the glycoprotein of the tumor cell is associated with the carbohydrate epitope of said glycoprotein of said tumor cell, and wherein the specific antibody is specific for the combination of the carbohydrate epitope and the peptide epitope of said glycoproteine.
  • the present invention provides a nucleic acid library encoding antibodies, wherein each of the nucleic acids in the library comprises
  • a second nucleic acid sequence selected from a repertoire of nucleic acids sequences comprising one or more nucleic acid sequences encoding a VL-domain that binds a peptide epitope of said glycoprotein of said tumor cell, for in-vitro identification of a specific antibody from said library which specifically binds a tumor cell, wherein the peptide epitope of the glycoprotein of the tumor cell is associated with the carbohydrate epitope of said glycoprotein of said tumor cell, and wherein the specific antibody is specific for the combination of the carbohydrate epitope and the peptide epitope of said glycoproteine.
  • nucleic acid libraries comprising a plurality of nucleic acid sequences, wherein each nucleic acid sequence of the plurality of nucleic acid sequences encodes an amino acid sequence forming at least a part of an antibody as described herein.
  • the present invention provides a nucleic acid library encoding a plurality of antibodies, wherein each of the plurality of nucleic acid sequences encoding antibodies comprises
  • a second nucleic acid sequence selected from a repertoire of nucleic acid sequences comprising one or more nucleic acid sequences encoding a VL-domain, that binds a peptide epitope of said glycoproteins of said tumor cells, for in-vitro identification of a specific antibody from said library which binds a tumor cell, wherein the peptide epitope of the glycoprotein of the tumor cells is associated with the carbohydrate epitope of said glycoprotein of said tumor cells, and wherein the specific antibody is specific for the combination of the carbohydrate epitope and the peptide epitope of the glycoprotein.
  • the nucleic acid library comprises in the range of 10 8 -10 9 nonidentical clones, such as at least 10 4 , 10 5 , 10 5 , 10 7 , 10 8 , 10 9 , or more non-identical nucleic acids.
  • first and the second nucleic acid sequences are linked by a nucleic sequence encoding a peptide linker connecting the encoded VH sequence with the encoded VL sequence.
  • the peptide linker is discussed above. Such linkers are generally known to the skilled artisan.
  • vector libraries comprising a nucleic acid library as described herein.
  • Examplary expression vectors for inserting nucleic acid libraries disclosed herein may comprise eukaryotic or prokaryotic expression vectors.
  • the nucleic acid library encoding antibodies are expressed using phage display technology.
  • cell libraries comprising a nucleic acid library as described herein.
  • the present invention provides a method for identifying an antibody for targeting a tumor cell, comprising preparing an antibody library as disclosed herein, and screening said library to identify one or more tumor targeting antibodies.
  • the antibody library is prepared as a phage displayed library, and the screening comprises biopanning of the antibody library using specific tumor glycopeptides or the intact glycoprotein.
  • the method further comprises isolating the tumor targeting specific antibody, and optionally purifying the antibody.
  • Antibody isolation and purification may be done by any common method recognized by a person skilled in the art.
  • the process for identifying and isolating a tumor-specific antibody comprises the steps
  • the identification of tumor cell specific antibody candiates comprises biopanning of the antibody library using a (specific) tumor glycopeptide or glycoprotein of said tumor cell, preferably O-glycosylated peptides or proteins, preferably where the glycosylation consists of short truncated O-glycan(s), such as a Tn-mucin or other O- glycosylated proteins having mucin-like motifs, as a purified peptide/protein or expressed on cell surfaces/tissues.
  • the antibody library may be a phage display library, yeast display library, ribosomal display library, or similar, as recognized by a person skilled in the art.
  • the antibody library is a phage display library.
  • the antibody library is a phage display library prepared by a method comprising the steps 1) mRNA isolation from a spleen (for the preparation of Tn and STn binding domains, the donor animal (e.g. mice) is immunized with a glycoprotein or glycopeptide comprising the short truncated O-glycans Tn or STn on the surface; for the preparation of the peptide binding domain, the the spleen is taken from a naive donor animal), 2) cDNA synthesis from said mRNA, 3a) amplification from said cDNA using a specific set of primers to obtain a first nucleic acid sequence encoding the VH domain, 3b) application from said cDNA using a mix of primers to obtain multiple nucleic acid sequences encoding the repertoire of VL domains, 4) assembly of the first nucleic acid sequence endocing the VH-domain and a second nucleic acid sequence from the multiple nucleic acid sequences encoding the V
  • such phage display library may be prepared as illustrated in Figure 2, comprising 1) mRNA isolation from mouse spleen. 2) cDNA synthesis with reverse transcriptase using random hexamers. 3) PCR amplification from cDNA template to obtain the VH domain using a specific set primers, and the repertoire of VL domains using a mix of VL primers. 4) PCR assembly of VL-domain repertoire and specific VH- domain using 5' phosphorylated outer primers.
  • Selections from the antibody library may be performed by several rounds of interogation (panning) with immobilized biotinylated target antigen (eg. bisTn-MUCl glycoprotein) on streptavidin-coated magnetic beads. After each round of selection, phages are eluted, amplified and precipitated. Removing extraneous phage antibodies by absorption against non-targets (negative binders), naked beads, plastics, proteins, peptides or normal human cells may also be performed as needed. Sequencing (NGS) of enriched phages after each round of panning provides a fingerprint of VL-domain antibody sequences corresponding to target antigen structure and peptide sequence.
  • biotinylated target antigen eg. bisTn-MUCl glycoprotein
  • Polyclonal phage ELISA may be used to confirm enrichments for target binder (e.g bis Tn-MUCl target protein/peptide). Phage pools may then be converted to soluble scFvs and expressed as individual scFvs. Expression of the scFvs in the supernatant may be assessed with dot blot analysis.
  • target binder e.g bis Tn-MUCl target protein/peptide
  • Screening for tumor-specific scFv clones in said phage library may be done using a ELISA binding assay, glycoprotein/peptide microarray and biolayer interferometry (OCTET) against target protein/peptide and control proteins/peptides (non targets), provided scFv antibodies targeting the selected glycoprotein antigen with high specificity and affinity.
  • OTET glycoprotein/peptide microarray and biolayer interferometry
  • Binding (FACS) of selected scFv to tumor cells expressing the target glycoprotein antigen such as breast adenocarcinoma cell lines MCF7, MDA-MD-231 COSMC KO may further be used to confirm tumor specificity.
  • the present invention provides a method as disclosed herein, wherein the antibody library is a phage displayed library, and the screening comprises biopanning of the antibody library using a specific tumor glycopeptide, such as Tn-MUCl, Tn-CD43, Tn-MUC4, Tn-MUC16, Tn-MUC13, etc.
  • a specific tumor glycopeptide such as Tn-MUCl, Tn-CD43, Tn-MUC4, Tn-MUC16, Tn-MUC13, etc.
  • Tn-tumor target are - as a non-limiting example - disclosed in the review by Kudelka et al 2015.
  • the present invention provides an antibody library and method for identifying specific antibodies against combined glycoside-peptide epitopes (combotobes) on specific glycoproteins, such as Tn, bis-Tn, STn or bis-STn epitopes on specific cancer cells.
  • specific glycoproteins such as Tn, bis-Tn, STn or bis-STn epitopes on specific cancer cells.
  • glycoside-peptide epitope binding antibodies which may have therapeutic effects due to their ability to specifically bind to specific glycoproteins on for example cancer cells.
  • the antibody library provided herein facilitates identification of an antibody that may be used to identify (diagnose) or treat a disease or disorder, such as cancer.
  • a proliferative disorder wherein the proliferative disorder is cancer
  • methods for treatment of a proliferative disorder comprising identifying and isolating anti-cancer cell antibodies by screening an antibody library of the present invention for antibodies having high specificity for said cancer cells, and administering to a subject diagnosed with said cancer disease the antibody identified as described herein.
  • a particular method of treatment involves the use of the combotope antibodies of the present invention in loading natural killer (NK) cells with the specific antibodies for targeting the NK-cells to the target cells, e.g. cancer cells.
  • NK-cells may be harvested from the patient to be treated prior to the loading with the antibodies or provided as donor NK-cells.
  • the loading may be in the form of the antibody or antibodies per se or as (a) nucleotide sequence(s) encoding the specific antibody/antibodies.
  • the specific antibodies are linked to a cell toxin or a non-toxic precursor thereof or a similar cytotoxic effector molecule of cell death.
  • the skilled artisar would readily know which effector molecules could be useful.
  • Further provided herein are methods for treatment of a proliferative disorder wherein the cancer is selected from lung, head and neck squamous cell, colorectal, melanoma, liver, classical Hodgkin lymphoma, kidney, gastric, cervical, merkel cell, B-cell lymphoma, and bladder cancer.
  • the cancer is a solid tumor.
  • aspects of the invention include administering any one of the specific antibodies identified as described herein to a subject identified as having aberrant/truncated O- glycosylation (e.g. crizated O-glycosylation of MUC1 protein) as compared to a reference level, (e.g. level in a non-cancerous cell).
  • a reference level e.g. level in a non-cancerous cell.
  • the truncated O- glysylation is selected from Tn and STn antigens, such as Tn-MUCl.
  • the present invention provides combotope antibodies as disclosed herein for use in treatment of a disease associated with aberrant/truncated O- glycosylation. In one embodiment, the present invention provides combotope antibodies as disclosed herein for use in treatment of a disease associated with Tn and/or STn antigens. In one preferred embodiment, the present invention provide combotope antibodies as disclosed herein for use in treatment of cancer.
  • the disclosure features methods that include administering any one of the specific antibodies identified as described herein, or a composition comprising such antibody, e.g. a cell composition, antibody-drug conjugate, or antibodyradioisotope conjugate) to a subject in need thereof, said subject having, or identified or diagnosed as having a cancer characterized by hypoglycosylation of peptide epitopes in the cancer cells (e.g., pancreatic cancer, epithelial cancer, breast cancer, colon cancer, lung cancer, ovarian cancer, or epithelial adenocarcinoma).
  • a cancer characterized by hypoglycosylation of peptide epitopes in the cancer cells
  • the libraries of the present invention comprise antibodies that are adapted to the species of an intended therapeutic target.
  • these methods include "mammalization".
  • the mammal is mouse, rat, equine, sheep, cow, primate (e.g., chimpanzee, baboon, gorilla, orangutan, monkey), dog, cat, pig, donkey, rabbit, and human.
  • the antibodies are intended for human therapeutic targets, and therefore humanized.
  • Tumor-specific antibodies are used in immuno-oncology to target cancer cells and activate the immune system to attack these cells. They can work by directly binding to cancer cells and triggering an immune response, or by targeting molecules on cancer cells that suppress the immune response. This can lead to increased tumor cell death and/or slower tumor growth.
  • Tumor-specific mAbs are often used in combination with other immune-based therapies, such as immune checkpoint inhibitors or CAR-T cell therapy, to enhance the anti-tumor immune response.
  • Tumor-specific monoclonal antibodies are also used in antibody-drug conjugates (ADCs) to deliver a cytotoxic drug directly to cancer cells.
  • ADCs antibody-drug conjugates
  • the mAb in the ADC is designed to recognize and bind to a specific protein on the surface of cancer cells, and once bound, the cytotoxic drug is released to kill the cancer cell.
  • the advantage of using an ADC is that it can selectively deliver the drug to cancer cells, minimizing the damage to healthy cells.
  • Some examples of ADCs that use tumor-specific mAbs include trastuzumab emtansine (T-DM1) for HER2-positive breast cancer and inotuzumab ozogamicin for acute lymphoblastic leukemia.
  • the antibodies of the present invention - i.e. antibodies identified using the antibody library of the present invention - are used to target cancer cells, such as to activate the immune system to attack the cancer cells.
  • the present invention provides an antibody as disclosed herein for use in treatment and/or prevention of cancer.
  • the cancer is lung, head and neck squamous cell, colorectal, melanoma, liver, classical Hodgkin lymphoma, kidney, gastric, cervical, merkel cell, B-cell lymphoma, or bladder cancer.
  • the cancer is a solid tumor.
  • the antibodies of the present invention are used in combination with other immune-based therapies, such as immune checkpoint inhibitors and/or CAR- T cell therapy, to enhance the anti-tumor immune response.
  • the antibodies of the present invention is used in antibody-drug conjugates (ADCs), such as to deliver a cytotoxic drug directly to cancer cells.
  • ADCs antibody-drug conjugates
  • the present invention provides a method of treating a cancer comprising administering a formulation comprising at least one specific antibody as disclosed herein to a patient in need thereof.
  • the antibody administered is conjugated to a cytotoxic moiety or loaded into a NK-cell, such as the patients own NK-cells, for being presented on the surface thereof.
  • Specific antibodies for use in treatment of cancer may be selected from a list of antibodies, wherein all antibodies comprise
  • the antibody administered to the patient in treatment of cancer comprises
  • VH domain comprising an amino acid sequence having at least 90% sequence homology to SEQ ID NO. 1, and amino acid residues H32, A33, H35, Y50, and S99 and/or S52, N55, and D57, with respect to SEQ ID NO. 1, and (ii) a VL domain comprising an amino acid sequence selected from SEQ ID NO. 9-21.
  • the antibody administered to the patient in treatment of cancer comprises
  • VH domain comprising an amino acid sequence having at least 90% sequence homology to SEQ ID NO. 28, and comprising amino acid residues (i), T28, T30, H32, A33, H35, Y50, S99, L101, A102 and L103, (ii), T28, T30, S52, N55, D57, L101, A102 and L103, or (iii) T28, T30, H32, A33, H35, Y50, S52, N55, D57, S99, L101, A102 and L103, with respect to SEQ ID NO. 28, and (II) a VL domain comprising an amino acid sequence selected from SEQ ID No. 22-23.
  • the antibody administered to the patient comprises (i) a first VH domain comprising an amino acid sequence having at least 90% sequence homology to SEQ ID NO. 1, and amino acid residues H32, A33, H35, Y50, S52, N55, D57, and S99, with respect to SEQ ID NO. 1, and (ii) a first VL domain comprising an amino acid sequence selected from SEQ ID NO. 9-21; or
  • a second VH domain comprising an amino acid sequence having at least 90% sequence homology to SEQ ID NO. 28, and comprising amino acid residues T28, T30, H32, A33, H35, Y50, S52, N55, D57, S99, L101, A102 and L103, with respect to SEQ ID NO. 28, and (II) a second VL domain comprising an amino acid sequence selected from SEQ ID No. 22-23.
  • the present invention concerns diagnostics.
  • a Tn- or STn- binding monoclonal antibody as disclosed herein can be utilized as a diagnostic tool for cancer in a subject by targeting the Tn/STn combotope found on the surface of many cancer cells but rarely present in normal cells.
  • the subject may be a human or an animal.
  • the process begins with the administration of the Tn- or STn-binding mAb, which has been designed to specifically recognize and bind to the Tn- or STn- antigen. Once administered, the mAb circulates through the body and binds to the Tn- or STn- antigens expressed on the surface of cancer cells.
  • This binding can be detected and visualized using various imaging techniques, such as PET, MRI, or fluorescence imaging, depending on the label attached to the mAb.
  • imaging techniques such as PET, MRI, or fluorescence imaging, depending on the label attached to the mAb.
  • the presence and distribution of the Tn/STn-antigen- mAb complexes in the body can then be analyzed to determine the presence, extent, and possibly the type of cancer.
  • This method offers a targeted approach to cancer diagnosis, potentially allowing for earlier detection and a more precise understanding of the cancer's location and spread, which is crucial for effective treatment planning.
  • Combotope antibodies of the present invention may be used in such diagnostics approach.
  • a tissue sample is collected from the patient suspected of suffering from a cancerous state, for example in form of a biopsy from the suspected cancerous tissue or by removal of whole or parts of the cancerous tissue for subsequent diagnosis ex vivo by binding one or more combotobe antibodies of the present invention to the tissue sample followed by identification of specific binding of the antibody by methods commonly known to the skilled artisan working in the field of identifying tissue antigen targets by way of immune detection.
  • Such diagnostic methods are generally known and performed on a daily basis in hospitals around the world.
  • the antibodies of the present invention are preferably humanized. This may be done is several different ways as acknowledged by a person skilled in the art.
  • a non-limiting example of such humanization of antibodies comprises the following steps:
  • the first step in humanizing a mouse monoclonal antibody is to identify an antibody with the desired specificity and affinity. This is typically done by screening a large library of mouse monoclonal antibodies using techniques such as ELISA or flow cytometry.
  • a human antibody framework is selected based on its structural similarity to the mouse antibody framework. This is important because it ensures that the humanized antibody retains the overall structure and stability of the original antibody.
  • the mouse-derived antigenbinding regions also known as complementarity-determining regions (CDRs) are replaced with human-derived CDRs while retaining the overall structure of the antibody. This is done using genetic engineering techniques such as PCR, cloning, and site-directed mutagenesis. Often a minor number of amino acids need to be replaced. Conservative substitions may be made without changing the properties of the antibody. The skilled artisan knows have to select conservative amino acids for replacement purposes in the antibody, for example to humanize the antibody or to facilitate the synthesis and production of the antibody. The substitution(s) are made by changing the nucleotide code(s) in the nucleotide sequence encoding the domains of the antibody.
  • the humanized antibody is tested for its specificity, affinity, and functionality. This is typically done using techniques such as ELISA, flow cytometry, and Western blotting.
  • the humanized antibody is also tested for its immunogenicity, which is its ability to trigger an immune response in humans. If the humanized antibody is found to be safe and effective, it can be further developed for use in human therapies. In case of immunogenicity of a certain promising combotope antibody, some of the amino acids may be substituted by concervative counterparts, however securing the the specificity and efficaty remains unchanged or even improved.
  • the present invention provides a method for identifying a glycopeptide target, said target comprising a Tn and/or STn epitope, such as a glycopeptide target of a cancer cell.
  • the library of the presnt invention is used for the identification of such glycopeptide targets by for example immunoprecipitation and mass spectrometry: This approach involves incubating the phage display antibody library with a cell lysate or tissue sample and allowing the antibody to bind to its target protein. The antibody-protein complex is then isolated by immunoprecipitation and subjected to mass spectrometry analysis to identify the protein.
  • Protein microarray Protein microarrays are arrays of immobilized antibodies that can be used to identify protein targets of antibodies. By incubating the phage display antibody library array with a cell lysate or tissue sample, and detecting binding, it is possible to identify the target protein.
  • the present invention discloses a method for identifying a glycopeptide target, said target comprising a Tn and/or STn epitope and a peptide target, such as a glycopeptide target on a cancer cell, said method comprising the steps of i) preparing an antibody library as disclosed herein, and ii) incubating said antibody library with a sample comprising the glycopeptide target, iii) analyzing one or more antibody-peptide complexes obtained from step ii) to identify the amino acid sequence of the peptide epitope of said glycopeptide target.
  • the sample is a cell or tissue sample, such as lysed cells or tissues.
  • Glycoproteins may be isolated from the lysed Academic and used for the screening.
  • the antibody library is preferably prepared as a phage display library, and the glycopeptides targets may be identified by analysis of the antibody-glycopeptide complexes using mass spectrometry analysis, or other similar method as recognized by a person skilled in the art.
  • the VH-domain of the combotope antibody of the present invention binds a carbohydrate epitope (Tn and/or STn) of a glycoprotein of the cancer cell, while the VL-domain of the combotope antibody binds a peptide epitope of the same glycoprotein associated with the carbohydrate epitope.
  • VH-domain may be characterized as a carbohydrate epitope binding VH domain, and further whether a VL-domain may be characterized as a peptide epitope binding VL domain.
  • amino acid sequence in question comprises amino acid residues positions corresponding to T28, T30, H32, A33, H35, Y50, S52, N55, D57, S99, L101, A102 and L103, with respect to SEQ ID NO. 28 - i.e. whether the sequence in question have the key residues needed for the VH domain functionality of STn binding.
  • At least three amino acids corresponding to the above mentioned amino acid residues need to be present in the VH domain for it to function as a VH domain binding Tn and/or STn.
  • the VH-domain does not contribute to or interfere with binding any peptide epitopes on the glycoprotein, which means that binding of the VH-domain to the carbohydrate epitope is not influenced, interfered or affected by any peptide epitopes on the glycoprotein. Only in this way, the VH-domain can be freely be combined with any VL-domain of choise for a "clean" binding without any disturbing binding between the VH-domain of the antibody and a peptide epitope, which unwanted binding may distort the result - in e.g. diagnosis or specific drug delivery, targeting a specific tumor glycoform (Tn or STn) on a given protein.
  • Tn or STn tumor glycoform
  • VL-domain For the VL-domain, this is discovered as disclosed herein, and will be unique to each target.
  • the identified VL domains have been (1) evaluated based on specificity and (2) correlated with other VL-domain sequences to identify common traits in the CDRs. X- ray may then confirm these traits.
  • MCF7, MDA- MD-231 WT and COMSC KO were maintained in DMEM+GlutaMax (Gibco, 32430-027) supplemented with 10% FBS (FisherScientific, 11550356), 1% penicillin-streptomycin (FischerScientific, 15140122) and 1 mM sodium pyruvate (Gibco, 11360).
  • FBS FisherScientific, 11550356
  • penicillin-streptomycin Factorific, 15140122
  • 1 mM sodium pyruvate Gibco, 11360.
  • Jurkat cells were maintained in RPMI (Life Technologies, 32404014) supplemented with 10% FBS, 1% penicillin-streptomycin and 2 mM L-glutamine (Sigma, G7513).
  • HEK293 cells were maintained in Freestyle media (Thermo Scientific, 15285885).
  • XLl-Blue electrocompetent cells were supplied by Agilent (Agilent, 200228). TGI for phage display were kindly provided by Peter Kristensen from Aalborg University. Vectors pAKlOO phagemid and pJB33 expression vector were kindly provided by Plunthum from University of Zurich, both with chloramphenicol antibiotic resistance. E. coli TGI and XL1- blue electrocompetent cells were cultures in 2xYT broth media. Liquid media was supplemented with 25 pg/mL chloramphenicol and 2% glucose unless otherwise stated.
  • GraphPad prism 9 was used for graph design and Biorender for image design.
  • CLC Main workbench 8.0 software was used for sequence alignment.
  • G2D11 is a mouse derived anti-Tn-scFv mAb.
  • ScFv consists of VH domain SEQ ID NO 1 and LV domain SEQ ID NO. 2 joined by pepide linker (GGGGS)4.
  • ScFv G2D11 crystals were prepared by the sitting drop technique and by using appropriate precipitant solutions. The resulting crystals were used to solve the structure at a resolution of 1.9 A and interpret the density map (Figure 3). Despite two molecules were present in the asymmetric unit and that contacted weakly between each other, analytical ultracentrifugation showed that this monomeric form behaved as a monomer either in the absence or presence of the bis-Tn-MUCl peptide APGS*T*AP where * denotes a GalNAc moiety (SEQ ID NO.: 55).
  • VL and VH The glycopeptide laid within a surface groove formed by the light (L) and heavy (H) chains (hereafter VL and VH, respectively), and in particular the two GalNAc moieties were recognized by residues from the three hypervariable regions of the VH ( Figure 3).
  • Thr-bound GalNAc was also intimately recognized by the scFv-G2Dll.
  • Ser52 H was engaged in hydrogen bond interactions with the carbonyl group, OH3 and OH4, while Asn55 and Asp57 side chains interacted with OH4.
  • G2D11 is more open and can easier tolerate binding any combination TnThr/TnSer, TnSer/TnThr, TnSer/TnSer, TnThr/TnThr.
  • G2D11 VH-domain was aligned with other known anti Tn-antibody VH-domains using CLUSTALW (using standard settings for multiple alignment parameters - i.e. scoring matrix: blosum62, gap opening penalty: 10, and gap extension penalty 0.2). Based on this alignment, conserved amino acid residues in the CDR1, CDR2 and CDR3 regions relevant for the functionality of the VH-chain (i.e. binding the GalNac) were identified, as illustrated in Figure 4 by the arrows.
  • amino acid residues H32, A33, and H35 in CDR1 amino acid residues Y50, S52, N55, and D57 in the CDR2, and also amino acid residue S99 in the CDR3 should preferably be conserved for the VH domain.
  • Example 2 Tn-template phage display libraries conceptualization Based on structural the VH-domain X-ray data and the observation that bisTn-binding was due to the VH-domain (as disclosed in Example 1), an antibody library was conceptualized, wherein each antibody comprised the VH chain of the previously identified scFv G2D11, providing recognition support for the glycoside part of the antigen, while the VL-domain was variable originating from naive mice, creating a scFv phage display library, termed as Tn-template library, which can be screened for a specific scFv, of which the VL-domain would provide recognition support to the underlying peptide antigen within the combotope.
  • RNAIater RNA stabilization reagent for cDNA synthesis with random hexamer primers (FisherScientific, 10609275) and Superscript IV Reverse Transcriptase (Invitrogen, 18090010).
  • the constant VH gene as well as the VL antibody specific genes were amplified by PCR using Q5 Hot Start High-Fidelity DNA Polymerase (NEB M0494S).
  • VL and VH genes were gel exctracted and assembled with 5' phosphorylated outer primers to allow the rolling circle amplification in the next step.
  • RCA improves restriction enzyme (Sfil) cutting of the scFv genes.
  • Sfil restriction enzyme
  • the assembled scFv fragments were sub-cloned in the Sfil-digested phagemid vector pAKlOO using Electroligase (NEB M0369) for 16 h at 16oC/25oC.
  • the phagemid pool with a variety of scFv fragments was electroporated in XLl-Blue electrocompetent cells (Agilent, 200228).
  • Bacteria were spun and pellet was resuspended in liquid media without glucose but supplemented with antibiotics and isopropyl B-D-l-thiogalactopyranoside (IPTG in 1: 1000 dilution) to induce phage production. Overnight cultures were centrifuged to remove bacteria pellet and phages were precipitated form the supernatant by adding ice cold PEG/NaCI (20% w/v PEG6000, 2.5 M NaCI) in 1:4 ratio. After 1 h incubation on ice, precipitated phages were spun at 10,800xg for 30 min followed by a centrifugation at 5,000xg for 5 min.
  • IPTG isopropyl B-D-l-thiogalactopyranoside
  • HCI hydrochloric acid
  • polyclonal phagemids with the different scFv fragments were purified with the GeneJet Plamsid Miniprep Kit (Thermo Fischer, K0503) according to the protocol, digested with Sfil restriction enzyme for 20 min at 50oC and ligated in the pJB33 expression vector using T4 electroligase for 1 h at 65oC.
  • the pool of the different constructs was electroporated in XLl-Blue electrocompetent cells, cells were recovered in SOC media, incubated for 1 h at 220 rpm at 37oC and then cells were spread on agar plates and incubated overnight at 37oC.
  • Bacteria were harvested at 6,000xg for 10 min and pellet was resuspended in ice-cold 100 mM Tris, 20% w/v sucrose solution with EDTA-free protease inhibitor cocktail (ThemroFischer, A32965), pH 8. After centrifugation at 8,000xg for 10 min, pellet was resuspended in ice-cold 5 mM MgSO4 in MQ solution. Pellet was centrifuged at 8,000xg for 10 min and the 2 fractions were pooled together and centrifuged at 12,000xg for 60 min to remove any cell debris.
  • polyclonal phage ELISA plates were incubated with polyclonal phages in serial dilutions in blocking buffer for 2h followed by incubation with secondary antibody incubation for 1 h. Bound phages were detected with mouse monoclonal anti M13- HRP antibody (Nordic Biosite 58-11973-MM05T-H-100) at 1: 10000 dilution.
  • scFv monoclonal scFv ELISA
  • antigen concentration was 50 nM. 50 uL of supernatant from the overnight culture of each clone was added per well.
  • antigen coating was at fixed concentration of 330 nM peptides and scFv were titrated 5- fold starting from 300 nM. Bound scFv were detected with a mouse monoclonal anti-His HRP (C-term) (Invitrogen 46-0707) at 1:2000 dilution.
  • the streptavidin biosensor tips were loaded with the biotinylated target glycopeptide in kinetics buffer for 300s, followed by an additional equilibration step of 100s. Association of scFvs in a range of different concentrations was performed for 300s. Finally, the dissociation was monitored with kinetics buffer for 300s. The association and dissociation responses were processed with the Octet Software (Version 12). Interferometry data was globally fitted to a 2: 1 model calculating the affinities and rate constants.
  • Sequencing was performed with Oxford Nanopore Technology (ONT). After each selection round, bacteria were scraped from the agar plate and an aliquot of the homogenous suspension was used for DNA purification using the GeneJet Miniprep Kit according to the manufacture's protocol. For the unselected libraries, homogenous suspension of scraped bacteria were used for DNA purification using Nucleobond Xtra EF Plasmid purification (MACHEREY-NAGEL GmbH & Co, 740422.50M) according to the manufacture's protocol. The set of primers that were used are found in the sequence listing (SEQ ID NOs. 50-53). Three pg of plasmid DNA were used as input material.
  • Nanopore sequencing and data analysis was performed according to Karst et al 2021 with the following modifcations: a 0.8 x volume of AMPure XP beads was used for DNA clean-up after early and late PCR, all the DNA washes for the purification were performed with 80% ethanol. DNA was quantified using the Qubit dsHS DNA assay (Thermo Fisher Scientific). After late PCR, a 1% agarose gel was performed to verify the correct product size. Samples prepared for the R9 flow cell the SQK-LSK110 ligation sequencing kit protocol was used while samples prepared for the RIO flow cell the SQK-LSK114 ligation sequencing kit protocol was used. Samples run on a flow cell and sequencing was performed on a MinlON MklB device for 72 h.
  • Example 3 MUC1 as a proof of concept
  • MUC1 was used as proof of concept for the constructed libraries to identify binders against MUC1. Comparing the newly identified scFvs and the known mAbs in terms of sequence and antibody activity, the effectiveness and the functionality of both libraries was determined. The identified scFv were sequenced followed by VL chain analysis and characterized for their specificity on ELISA, cell binding assays and kinetic studies with BLI.
  • Tn template library wherein each antibody in the library comprises G2D11 VH domain (SEQ ID NO. 1) was subjected to three rounds of selections by immobilizing target peptide 1 (see Table 1) on streptavidin-coated beads using target antigens. After each round of selection, phages were eluted, amplified and precipitated. Phage stocks after each round were titrated and analysed on polyclonal phage ELISA ( Figure 5). Polyclonal phage ELISA showed specific binder enrichment for MUC1 target peptide in every round with zero non-specific binders against streptavidin.
  • the control peptide that was used in this study is IgA (see Table 1) that is produced in mucosal membranes and plays a significant role in their immunity. It has N- and Clunked glycosylation sites and is involved in a number of pathological conditions such as IgA deficiency and IgA nephropathy. Clones that showed no reactivity against the control peptides were selected for further characterization. In addition, sequences of the sixty one clones with sanger sequencing were obtained and alignment with VL sequences of 5E5 and 2D9 showed the differences in the CDR of the VL chains. Key binding features as presented in Table 2 were also present in VL-sequences from 5E5 and 2D9 further corborating their importance for interactions with the peptide backbone and determination of specificities.
  • the selected clones were produced in larger scale for evaluation in ELISA, kinetic studies and cell binding assays.
  • the scFvs were expressed and purified by His- tag affinity purification. Purity of the scFvs was checked on SDS-PAGE Coomassie analysis and western blot to confirm the presence of His-tag.
  • soluble scFvs were screened for their binding at fixed concentration of MUC1 target peptide 1 and control peptide 8 (see table 1) Results are found in Figure 8A and 8B.
  • scFvs were screened for cross-reactivity on other MUC1 peptides 2, 3, 4 and 5 (see Table 1). Results are found in Figure 9.
  • Negative binding on unglycosylated MUC1 confirmed the library hypothesis that scFvs against the peptide backbone only cannot be selected.
  • scFv D5 showed binding to monoTn-MUCl glycopeptide for concentration > 10 nM.
  • clone D5 bind to peptide 3, while clones H3 and D3 also bind to peptide 3, but only at high concentrantion.
  • the clones that showed higher specificity for the target peptide 1 and zero crossreactivity to IgAl hinge region were chosen for further assessment.
  • A3 and D2 clone demonstrated no binding to IgAl hinge region while D3 and H3 demonstrated binding at high antibody concentrations.
  • the rest of the scFvs recognized also IgAl in addition to MUC1. Based on scFv specificity on titration ELISA A3, D2 and D3 scFvs were chosen for biological evaluation and kinetic studies.
  • COSMC KO means a knock-out of the COSMC gene which is a chaperon required to help catalyze the transfer of Galbl-3 to the penultimale sugat Tn (GalNAc), generating the Tn-antigen and subsequent elongation.
  • this COSMC is a frequent phenomenon and the results of exposure of Tn and STn-antigens on tumor proteins.
  • mAb HMFG2 (anti-MUCl) was used as a positive control to confirm MUC1 expression on cells (data not shown) and 5E5 was used as control mAb to confirm the Tn glycoform presence on MUC1.
  • 5E5 was used as control mAb to confirm the Tn glycoform presence on MUC1.
  • All three scFv A3, D2 and D3 showed positive binding on MDA-MD-231 COSMC KO cells and negative binding on WT cells ( Figure 8C).
  • no positive binding was detected on MCF7 cells either for the selected MUC1 scFv clones or for 5E5 mAb control ( Figure 10A).
  • Neuraminidase treatment did not enhance binding of 5E5 control and selected MUC1 scFv clonesD3, D2 and A4 on MCF7 cells (data not shown).
  • the binding affinity (KD) of G2D11 and the newly identified MUC1 scFv clones were determined with BLI.
  • the biotinylated target peptide 1 was immobilized on streptavidin sensor tips. The only clones that showed higher binding affinity compared to G2D11 was clone D3 (Table 3).
  • VH-G2D11 binds bisTn in two orientations which makes it also more flexible in binding to Tn-structures in contrast to e.g. 5E5 that prefer a mono-Tn attached to a threonine (Thr) and less bis-Tn structures.
  • scFvs D3, D5, H3 shared the MUC1 specific binding motif YSY in CDR3 like in 5E5 which is a requirement for peptide backbone binding interactions as it has been showed previously with crystallography ( Figure 11). More specifically, Y98 L and Y100 L are contributing to the peptide binding.
  • scFvs A3 and D2 shared the motif WNY while scFvs B5 and A2 had the motif SSY (Table 4). In all scFvs as well as the 5E5 mAb the Y100 L is conserved.
  • CDR1 and CDR2 has some variations that can be linked to the target peptide sequence and size (e.g W50 L ) residue in cloe proximity to the peptide backbone.
  • the obtained data correlate with current known information (x-ray and interacting residues, with specificities) from reference mAb 5E5. It demonstrates that the present library concept can enrich and select sequences with same features as obtained with immunized mice and hybridoma targeting the same antigen. This is a major step forward and with an animal free rapid system. The approach also provides a very large number of additional clones for evaluation with similar or different VL-sequence oprions that potentially could be better or different binders. The methods provides a relative (to hybridoma) controlled and systematic approach identifying large numbers of candidates for evaluation.
  • Glycopeptides from Table 5 were printed on microarray chip, and the binding to these petides by scFv D3 and scFv A4 was compared with binding by scFv 5E5, scFv 2D9Chi, and scFv G2D11.
  • 2D9Chi comprises the 2D9 VL domain and G2D11 VH domain. Results are presented in Figure 12 (for simplicity, the heat map shows amino acids 9-19 of the peptides in Table 5).
  • ScFv 2D9Chi binds to two adjacent Tn antigens only in Ser-Thr sequence, contrary to scFV G2D11 which also binds two adjacent Tn antigens in the Thr-Ser sequence.
  • ScFv A3 showed the same epitope recognition as 2D9Chi. Exchange of Pro residue with Ala, abolishes the binding of all scFvs against the glycopeptide.
  • Example 4 Concept evaluation - CD43 as first example
  • CD43 was chosen as the first target to identify binders following the same procedure.
  • Target peptide 9 (see table 1) was immobilized on streptavidin beads and three selection rounds were performed. To determine efficient phage selection, polyclonal phage ELISA and nanopore sequencing took place. Both polyclonal phage ELISA ( Figure 13) and sequencing confirmed phage and sequence enrichments between the rounds. However, in the case of CD43 the sequences were grouped based on the combinations of CDR1, CDR2 and CDR3 as it is not evident which CDR is responsible for peptide backbone binding as in the case of MUC1 (Table 6).
  • the ten clones (named ori, Hl, -Al, F4, C5, A7, D3, G3, D7, H2) were tested for their binding specificity on titration ELISA as a first step. These ten clones are represented by SEQ ID NOs. 12-21 in the sequences listing. Only their VL-domain is listed for the scFv. VH domain is same of the scFV of all clones - i.e. the VH of G2D11 (SEQ ID NO 1). The VL and VH are joined by the peptide linker (GGGS)s.
  • GGGS peptide linker
  • D3 with the L99 L showed cross-binding to IgA hinge region peptide contrary to D7 and Hl that showed specific interaction with CD43 peptide and no cross-reactivity with TnlgA hinge or TnMUCl.
  • scFv C5 that has a Y99 L also cross reacted with the control peptide.
  • a second target example was provided with a different peptide sequence but still mucin-like (amino acid features for O-glycosylation, high content of S, T, P, etc.), and enrichment of VL-domain sequences with a unique CDR fingerprint especially for CDR1 and CDR3 was demonstrated, and also confirmed by x-ray.
  • the additional CDR1 interactions further increase specificity as demonstrated with ELISA (no cross-reactivity to TnlgA or TnMUCl).
  • ELISA no cross-reactivity to TnlgA or TnMUCl.
  • G2D11 VH-domain mutants were prepared based on the above identified amino acid positions potentially relevant for STn specificity:
  • M2 (LAL-TFT): I28T, A30T, P101L, de/G102, T103A and F104L;
  • M3 (LAL-TFT-G): I28T, A30T, D56G, P101L, de/G102, T103A and F104L;
  • VH residues are required VH residues to have binding towards bisSTn O-glycans are T28, T30, H32, A33, H35, Y50, S52, N55, D57, S99, L101, A102 and L103, with reference to SEQ ID NO 28.
  • Table 9 and Table 10 provides the top ten most enriched combinations and their percentages in every round for bisSTn-MUCl and monoSTn-MUCl, respectively. In both selections, specific MUC1 binding motif Tyr-X-Tyr can be identified. The VL-domain sequences confirm that the STn-library can be used, and generates similar data/clones as obtained for TnMUCl.
  • VL-domain sequences contains the same features as seen for Tn-MUCl and further consolidate the fingerprint related to MUC1 peptide target.
  • scFv ELISA specificity and VL sequence comparison three clones were selected for purification. These three clones (named C4, D3, C7) are represented by SEQ ID NOs. 22-24 in the sequences listing. Only their VL-domain is listed for the scFv. VH domain is same for the scFV of all clones - i.e. the mutated VH chain of G2D11 (SEQ ID NO. 28) (i.e. the G2D11 VH domain comprising the IFA to TFT mutation in CDR1 and PGTF to LAL in CDR3, as disclosed in example 5). The VL and VH are joined by the peptide linker (GGGS)s.
  • GGGS peptide linker
  • Table 11 summarizes the kinetic affinities for the MUC1 and CD43 specific scFvs indentified using the antibody libray according to the present invention.
  • Example 8 mono Tn/STn scFv binders
  • the antibody library concept of the present invention can be used to target monoTn/STn-peptide binders and not only bisTn/Stn-peptide binders, and this significantly increases utility as Tn/STn are situated as orfan, bis or in larger clusters.
  • a humanized scFv is generated by humanising the VL and VH immunoglobulin domains derived from the murine-originated antiCD43. Humanisation of VL and VH is performed in scFv format as follows:
  • CDRs complementarity-determining regions
  • Human V and J gene segments are chosen as template frameworks based on their identity to antiCD43 sequence, in-house analysis of individual and pairing frequency of V genes and previous experience of the use of particular templates for legacy humanisation.
  • the chosen human V gene frameworks are compared to the respective murine VH and VL sequences to identify potential sites that could undergo back-mutation to the corresponding mouse amino acid at that position.
  • In-house collated evidences rules for the importance of certain framework positions in the likely maintenance of CDR conformation (and antigen binding affinity) are used to identify back-mutations considered most significant (primary mutations) and those of lower significance (secondary mutations).
  • the extent of spatial clustering of the identified back-mutations is examined by analysing the crystallized molecular structure of mouse antiCD3.
  • Initial humanised VH and VL sequences are generated by constructing a straight graft of the mouse CDRs into the chosen human germline templates.
  • the apparent spatial clustering of back-mutation sites is used to reduce the potential number of variants of back- mutation containing humanised chains by introducing spatially-clustered mutations simultaneously.
  • scFv sequences comprised a (G4S)4 linker between VL and VH chains, and a C-terminal exa-His tag. scFv protein sequences are reverse translated and codon optimised.
  • All codon optimised DNA sequences are modified to include 5' and 3' adaptors suitable for HiFi cloning in the pET22b (+) and synthesised.
  • DNA sequences are synthesised by TwistBioscience as double-stranded fragments (gBIocks).
  • pET22b (+) backbone is linearized by PCR and the product is treated with Dnpl and cleaned with Monarch DNA&PCR cleanup.
  • the gBIocks is inserted in pET22b (+) using the NEBuilder® HiFi DNA Assembly Cloning Kit.
  • Ligation mixtures are transformed into DH5a competent E. coli and positive transformants selected on plates of LB agar supplemented with 100 ug/ml carbenicillin. Colonies for putative clones are cultured, plasmid DNA extracted, and DNA subjected to Sanger sequencing to identify correct clones.
  • Transient transfection of scFv-encoding construct into HEK 2936E suspension culture 250 pg of DNA for each plasmid construct is transfected using 293 Fectin transfection reagent into separate 250 ml cultures of HEK 293 6E cells (at a viable cell density of 1.85x10® cell/ml. The cultures are placed into a shaking 37°C incubator at 124rpm with 5% CO2. At 48 and 72 hours the cultures are supplemented with 6.2 ml tryptone (200 g/l) and 6.2 ml 3M fructose respectively.
  • the viability (%) and viable cell density (cell/ml) of each culture are measured every 24 hours using a Vi-Cell cell counter and viability analyser (Beckman Coulter). Once the cultures have reached ⁇ 70% viability, the cultures are harvested via centrifugation at 4415xg for thirty minutes at 4°C and filtered via 0.22 pm Millipore filter. The filtered supernatants are stored at 4°C until required for protein purification.
  • the humanized antibody is purified using Single- Step Affinity Protein Purification.
  • the scFv proteins are purified from the resulting supernatants via AKTA Express system (AKTA).
  • the supernatant is loaded at 5 ml/minute onto a 5 ml HisTrap Excel column pre-equilibrated with Buffer A (50 mM HEPES pH 7.5, 400mM NaCI, 20mM Imidazole). Once loaded, the column is washed in two column volumes of Buffer A at 5ml/minute back to baseline.
  • the proteins are eluted in a step elution of 50% Buffer B (50 mM HEPES pH 7.5, 00 mM NaCI, IM Imidazole).
  • the column is held in three column volumes of 50% Buffer B. During this step elution 0.5 ml fractions are collected in the purification of 88A, then 1 ml fractions are collected in all subsequent purifications. The elution step is continued until returned to baseline followed by a washout step at three column volumes of 100% Buffer B.
  • a single peak is expected at 280 nM on the resulting chromatogram indicating the elution of the protein of interest.
  • the fractions corresponding to this peak are pooled and transferred to a 5000 MW cut-off centrifugal concentrator.
  • the sample is buffer exchanged from Buffer B into 60ml PBS to separate the purified protein from the imidazole present in Buffer B.
  • the samples are concentrated down to ⁇ 1 ml.
  • the concentration is determined via nanodrop and the purified protein is diluted in PBS to obtain a final concentration of 1 mg/ml.
  • the final protein product is aliquoted and stored at -80°C for future use.
  • the objective of this example was to prepare a humanized clone of the mouse antibody scFv clone D3 (i.e. SEQ ID NO. 9 (VL domain) and SEQ ID NO. 1 (VH domain) joined by the peptide linker (GGGS)4).
  • a homology model of the parental mouse D3 Fv was built using Protein DataBank ID 2GKI, which provided the best template for VH/VL combined.
  • CDR loops were modeled, using a knowledge based approach resulting in 3 different models, each with other CDR templates.
  • the parental mouse D3 sequence was used to interrogate the human germline sequence database with the BLASTP functionality integrated within Bioluminate. Additionally, maturated human antibody sequences were queried too. This resulted in 2 germline sequences for the VH: IGHV1-3 and IGHV1-69 and 1 germline sequence for the VL: IGKV4-1. Maturated sequences found were:
  • Template PDB ID 7LKB Crystal structure of PfCSP peptide 21 with vaccine- elicited human anti-malaria antibody m42.127 with 95.9% match human germline.
  • Template PDB ID 7PS3 Crystal structure of antibody Beta-32 Fab with 94.9% match human germline.
  • the templates were used to directly graft the CDRs based on IMGT definition with guidance of bioluminate. The resulting sequences were used to build homology models for future inspection. 5. The template sequences with grafted CDRs were aligned and residue by residue analyzed on potential (structural) issues.
  • VH and VL sequences are listed separate, to allow testing of different combinations - i.e. different combinations of D3VHx + D3VLx. Additionally, scFv sequences D3LxHx, with Gly-Ser linker were provided. Please note that signal peptide(s) are not provided.
  • D3VHx + D3VLx were found of particular interest: D3VL1 + D3VH1, D3VL1 + D3VH2, D3VL2+D3VH3, D3VL3 + D3VH4, and D3VL4 + D3VH5.
  • Kd meassurements was obtained by using Streptavidin coated Tip-sensors (SAX) from Startorius using manufacturer conditions and reagent kit, see Figure 22. The results are presented in table 12.
  • Example 11 Binding of mouse antibody scFv clone D3 to different Tn-peptide targets
  • An antibody library for in-vitro identification of a specific antibody which binds a tumor cell wherein each antibody in said library comprises
  • a second antibody domain selected from a repertoire of second antibody domains, wherein the repertoire of second antibody domains comprises one or more second antibody domains which binds a peptide epitope of said glycoprotein of said tumor cell, wherein said specific antibody is specific for a combination of said carbohydrate epitope and said peptide epitope of said glycoprotein.
  • PREFERRED EMBODIMENT 2 The antibody library according to PREFERRED EMBODIMENT 1, wherein the carbohydrate epitope is covalently linked to the peptide epitope.
  • PREFERRED EMBODIMENT 3 The antibody library according to PREFERRED EMBODIMENT 1 or 2, wherein the first antibody domain is a VH-domain, and wherein the second antibody domain is a VL-domain.
  • PREFERRED EMBODIMENT 4 The antibody library according to any one of PREFERRED EMBODIMENTS 1-3, wherein the first antibody domain is a VH-domain, wherein the second antibody domain is a VL-domain, and wherein the antibodies in the library are scFv wherein the VH-domain is linked to the VL-domain via a peptide linker.
  • PREFERRED EMBODIMENT 5 The antibody library according to any one of PREFERRED EMBODIMENTS 1-4, wherein the carbohydrate epitope is selected from mono-Tn, bis-Tn, mono-STn, bis-STn, and a combination of mono-Tn and mono-STn.
  • VH domain is a mono- or bis-Tn-binding VH domain
  • VH- domain comprises an amino acid sequence having at least 90% sequence homology to SEQ ID NO. 1 and comprising amino acid residues H32, A33, H35, Y50, S52, N55, D57, and S99, with respect to SEQ ID NO. 1.
  • PREFERRED EMBODIMENT 6 The antibody library according to any one of PREFERRED EMBODIMENTS 1-5, wherein the first antibody domain is a mono- or bis- Tn-binding VH domain, and wherein said VH-domain comprises an amino acid sequence having at least 90% sequence homology to SEQ ID NO. 1 and comprising amino acid residues H32, A33, H35, Y50, and S99 and/or amino acid residues S52, N55, and D57, with respect to SEQ ID NO. 1.
  • VH domain is a mono- or bis-STn-binding VH domain
  • VH-domain comprises an amino acid sequence having at least 90% sequence homology to SEQ ID NO. 28 and comprising amino acid residues T28, T30, H32, A33, H35, Y50, S52, N55, D57, S99, L101, A102 and L103, with respect to SEQ ID NO. 28.
  • PREFERRED EMBODIMENT 7 The antibody library according to any one of PREFERRED EMBODIMENTS 1-5, wherein the first antibody domain is a mono- or bis- STn-binding VH domain, and wherein said VH-domain comprises an amino acid sequence having at least 90% sequence homology to SEQ ID NO.
  • PREFERRED EMBODIMENT 8 The antibody library according to any one of PREFERRED EMBODIMENTS 1-7, wherein the first antibody domain does not contribute to or interfere with binding any peptide epitope.
  • PREFERRED EMBODIMENT 9 The antibody library according to any one of PREFERRED EMBODIMENTS 1-8, wherein the repertoire of second anytibody domains is generated from a naive immune repertoire of VL-domains, an immunized immune repertoire of VL-domains, or a synthetically produced repertoire of VL-domains; preferably a naive immune repertoire of VL-domains from an animal, such as a mouse or human.
  • PREFERRED EMBODIMENT 10 The antibody library according to any one of PREFERRED EMBODIMENTS 1-9, wherein the antibody library is a phage display library.
  • PREFERRED EMBODIMENT 11 A method for identifying an antibody for targeting a tumor cell, comprising the steps of i) preparing an antibody library according to any one of PREFERRED EMBODIMENTS 1- 10, and ii) screening said library to identify one or more tumor targeting antibodies.
  • a method for identifying an antibody for targeting a glycoprotein comprising the steps of i) preparing an antibody library according to any one of PREFERRED EMBODIMENTS 1- 10, and ii) screening said library to identify one or more tumor targeting antibodies.
  • PREFERRED EMBODIMENT 13 The method according to PREFERRED EMBODIMENTS 11 or 12, comprising biopanning of the antibody library using a glycopeptide or glycoprotein of said tumor cell, prefarebly an O-glycosylated peptide or protein, such as Tn-mucin or other O-glycosylated protein having mucin-like motif, wherein said glycopeptide or glycoproteain is used in prurified form or expressed on a cell surface or tissue.
  • O-glycosylated peptide or protein such as Tn-mucin or other O-glycosylated protein having mucin-like motif
  • the antibody library is a phage display library prepared by a method comprising the steps 1) mRNA isolation from a spleen, 2) cDNA synthesis from said mRNA, 3a) amplification from said cDNA using a specific set of primers to obtain a first nucleic acid sequence encoding the VH domain, 3b) application from said cDNA using a mix of primers to obtain multiple nucleic acid sequences encoding the repertoire of VL domains, 4) assembly of the first nucleic acid sequence endocing the VH-domain and a second nucleic acid sequence from the multiple nucleic acid sequences ecoding the VL-domain repertoire, to form a joint contruct, 5) insertion of the construct into a phagemid vector, 6) insertion of the phagemid vector comprising the construct into E. coli to produce a bacterial library, 7) using the bacterial
  • PREFERRED EMBODIMENT 15 A method for identifying a glycopeptide target, said target comprising a Tn and/or STn epitope and a peptide epitope, such as a glycopeptide target of a cancer cell, said method comprising the steps of i) preparing an antibody library according to any one of PREFERRED EMBODIMENTS 1- 10, ii) incubating the antibody library with a sample comprising the glycopeptide target, and iii) analyzing one or more antibody-peptide complexes obtained from step ii) to identify the amino acid sequence of the peptide epitope of said glycopeptide target.
  • a specific tumor cell binding antibody comprising
  • PREFERRED EMBODIMENT 17 The antibody according to PREFERRED EMBODIMENT 16, wherein the VH domain comprises
  • PREFERRED EMBODIMENT 18 The antibody according to PREFERRED EMBODIMENT 16, wherein the VH domain comprises
  • PREFERRED EMBODIMENT 19 The antibody according to any one of PREFERRED EMBODIMENTS 16-18, wherein the VL domain comprises an amino acid sequence selected from SEQ ID NO. 9-24.
  • PREFERRED EMBODIMENT 20 The antibody according to any one of PREFERRED EMBODIMENTS 16-19, comprising a first and a second antigen-binding fragement comprising a first VH and VL domain and a second VH and VL domain, respectively, wherein the first VH domain binds a Tn-carbohydrate epitope and the second VH domain binds a STn carbohydrate epitope.
  • PREFERRED EMBODIMENT 21 The antibody according to any one of PREFERRED EMBODIMENT 20, wherein the first VH domain comprises a first amino acid sequence having at least 90% sequence homology to SEQ ID NO. 1, wherein the first amino acid sequence comprises amino acid residues H32, A33, H35, Y50, and S99 and/or amino acid residues S52, N55, and D57, with respect to SEQ ID NO. 1; and wherein the second VH domain comprises a second amino acid sequence having at least 90% sequence homology to SEQ ID NO.
  • the second amino acid sequence comprises amino acid residues (i), T28, T30, H32, A33, H35, Y50, S99, L101, A102 and L103, (ii), T28, T30, S52, N55, D57, L101, A102 and L103, or (iii) T28, T30, H32, A33, H35, Y50, S52, N55, D57, S99, L101, A102 and L103, with respect to SEQ ID NO. 28.
  • PREFERRED EMBODIMENT 22 The antibody according to any one of PREFERRED EMBODIMENTS 16-21 for use in treatment and/or prevention of cancer.
  • PREFERRED EMBODIMENT 23 A method of treating a cancer in a subject comprising administering a formulation comprising at least one antibody according to any one of PREFERRED EMBODIMENTS 16-21 to a patient in need thereof.
  • PREFERRED EMBODIMENT 24 The antibody according to any one of PREFERRED EMBODIMENTS 16-21 for use in diagnosing a cancerous condition in a subject.
  • PREFERRED EMBODIMENT 25 A method of diagnosing a cancerous condition in a subject comprising administering a formulation comprising at least one antibody according to any one of PREFERRED EMBODIMENTS 16-21 to the subject, and detecting the presence of an antigen-antibody complex comprising said at least one antibody according to any one of PREFERRED EMBODIMENTS 16-21 and a Tn- and/or STn- antigen.

Landscapes

  • Chemical & Material Sciences (AREA)
  • Health & Medical Sciences (AREA)
  • Organic Chemistry (AREA)
  • Immunology (AREA)
  • Life Sciences & Earth Sciences (AREA)
  • Biophysics (AREA)
  • Biochemistry (AREA)
  • General Health & Medical Sciences (AREA)
  • Genetics & Genomics (AREA)
  • Medicinal Chemistry (AREA)
  • Molecular Biology (AREA)
  • Proteomics, Peptides & Aminoacids (AREA)
  • Cell Biology (AREA)
  • Virology (AREA)
  • Peptides Or Proteins (AREA)
  • Medicines Containing Antibodies Or Antigens For Use As Internal Diagnostic Agents (AREA)

Abstract

The invention provides specifc Tn- and STn antibody libraries and methods for identifying specifc antibodies, which target Tn- and/or STn- glycosylation sites of any glycoprotein of choice, especially glycoproteins of cancer cell targets. The invention further provides antibodies identified by the new concept prosposed herein, which have combined specificity towards both the carbohydrate epitope as well as the peptide epitope of the glycoprotein.

Description

TITLE: Combotope antibody libraries
FIELD OF THE INVENTION
The invention concerns antibodies, antibody libraries and methods for identifying antibodies, which target Tn- and STn- glycosylation site of any protein site of choice, especially relevant for binding to cancer cell targets. The antiboides identified by the new concept proposed herein have specificity towards both the sugar epitope as well as the peptide backbone in a glycoprotein associated with or carrying the sugar epitope on a cancer cell.
The invention provides specific Tn- and STn antibody libraries and methods for identifying specific antibodies, which target Tn- or STn- glycosylation sites of any glycoprotein of choice, especially relevant for binding to cancer cell targets, such as tumor cells. The invention further provides antibodies identified by the new concept proposed herein, which have combined specificity towards both the carbohydrate epitope as well as the peptide epitope of the glycoprotein.
BACKGROUND
A dense layer of complex carbohydrate structures covers almost all eukaryotic cells. Tumor cells, contrary to their healthy counterparts, exhibit altered glycosylation patterns on the cell surface. Such altered cancer glycosylation includes increased sialylation, fucosylation, short truncated O-glycans and increased N-branching.
One of the key features of glycosylation changes in cancer is the short truncated O- glycans, the so-called tumor-associated carbohydrate antigens (TACAs): Tn and T antigens and the sialylated forms of them, STn and ST, respectively.
The Tn and STn are not commonly observed in any normal human or rodent tissues, but are highly expressed on many solid tumors/carcinomas. Thus, Tn and STn represent major targets of potential immunotherapy as well as being useful in diagnostic of cancereous states.
There are a number of cell surface glycoproteins with altered glycosylation that play key roles in cancer including MUC1, MUC4 and MUC6 from the mucin family along with CD43 and CD44 as examples.
MUC1 is the most well-studied mucin from the mucin family and is present in many adenocarcinomas displaying short truncated O-glycans. Under healthy conditions, MUC1 peptide core is heavily glycosylated and therefore masked by the O-glycan moieties that protect MUC1 from proteolytic cleavage enzymes. In adenocarcinomas, MUC1 proteins have shorter and less dense O-glycan side chains, resulting in exposure of the core domains of the protein on the cell surface. This altered glycosylation on MUC1 results in exposure of the epitopes MUCl-Tn and MUCl-STn to the immune system.
CD43 (leukosialin) is a type I transmembrane sialoglycoprotein that is abundant in hematopoietic cells, including lymphocytes, monocytes, granulocytes, natural killer cells, platelets except resting mature B cells, and erythrocytes. The human CD43 protein has a mucin-type extracellular domain rich in serine and threonine residues enabling extensive O-GalNAc glycosylation with significant molecular weight heterogeneity. CD43 glycoforms have been reported in several hematological and non-hematopoietic cancers, including the lung, breast, colon, cervix, and prostate, which express CD43 mostly in the early stages of tumor progression.
Known antibodies releated to Tn/STn include 5E5 (Macias-Leon et al 2020; Tarp et al 2007; Blixt et al 2010), anti-CD43 (Blixt et al 2012), 2D9 (Sorensen et al 2006; Tarp et al 2007; Blixt et al 2010), and 5F7 (US11161911B2). Further known antibodies releating to Tn/STn include G2D11 (Persson et al 2017), 3F1 (Kjeldsen et al 1988), 83D4 (Oppezzo et al 2004), 15G9 (Mazal et al 2013), 1E3 (Li et al 2009), and MLS128 (Yuasa et al 2012); 16E12.1D9.1B11 (WO 2023/034569 Al), and 1 A5-2C9 (US 2022/057402 Al), however, these antibodies are Tn-hapten binders with no or unknown binding contribution to the peptide/protein carrier, as will be further discussed herein.
Monoclonal antibodies to the Tn and STn antigens are notably difficult to generate and are expensive to produce, and their specificities are often not well characterized, especially in regard to whether these antibodies simultaneously recognize the Tn/STn and the protein backbone/carrier, a requirement for superior specificity and therapeutic use.
The development of therapeutic antibodies can be achieved via several strategies. Current technologies, such as the well-established method of hybridoma technology, rely on animal immunization, where the animals are challenged with the antigen of interest to elicit an immune response. Then, B cells from the spleen of immunized animals (mice, rat and rabbit) are isolated and the subsequent fusion with myeloma cells to produce hybridoma clones for testing. It is a very tedious, laborious, timeconsuming method and many times the fusion efficiency is low. Also, the effectiveness of the method relies largely on the immunogenicity of the antigen, which some antigen shown to be poorly immunogenic. Phage display facilitates expression of proteins on the surface of phages. It is a molecular technique using the filamentous phage, where foreign DNA, encoding a peptide, is inserted in nonlytic philamentous phage genome and expressed as a fusion protein together with the phage coat protein without affecting the phage infectivity. Phage display allows a large repertoire of antibodies or parts thereof to be displayed on phages for selection of high affinity binders against a target of interest.
SUMMARY OF THE INVENTION
The present invention provides a new antibody concept technology for rapid development of combotope antibodies (Abs) targeting Tn- and STn- glycosylation site of any glycoprotein site of choice, paving the way for a new generation of therapeutic opportunities in cancer treatment and diagnostics.
A novel structural and biochemical explanation, provided herein for the first time, of Tn- and STn- recognition specifically by the VH hypervariable region of the antibody, is combined with phage display screening for specific peptide recognition by the VL domains, thereby providing combotope antibodies with high speficicity and affinity to desired biological glycoprotein targets due to the combined recognition of the Tn and/or STn carbohydrate epitope and the protein backbone/carrier associated with the Tn and/or STn carbohydrate epitope.
Phage display is a fast method for antibody development. The constructed Tn and STn template libraries define the VH part of the antibody binding the desired glycoform, and the VL diversity will determine the peptide backbone specificity. As disclosed herein in the examples section, MUC1 was used as a proof of concept target for both libraries (Tn and STn) to evaluate the functionality of these two libraries. Single chain variant fragment (scFv)s with sequences being the same as or similar to the already known MUC1 antibodies were isolated from biopanning of the libraries. CD43 was further the first target that was used to identify scFvs against Tn-CD43 peptide. Both Tn-MUCl and Tn-CD43 scFvs were characterized for their specificity in vitro, and the scFvs demonstrated their potential uses as therapeutics on cancer cell lines. Furthermore, STn-MUCl scFvs are identidifed.
In that way, two template libraries have been constructed, a Tn-template library (see examples 2-4) and a STn-template library (see examples 5-6), where the VH domain pre-selects for the desired glycoform specificity, and the VL domain will determine the peptide backbone specificity. The constructed libraries are the first libraries to be specifically designed for glycosylated targets. The constructed libraries are the first libraries to be specifically designed for specific glycosylated targets defined by the specific glycoprotein bearing the short truncated O-glycans (Tn or STn) presented on the surface of many cancer cells. The terms "tumor" and "cancer" are used interchangeable throughout this disclosure. Any differences between the meaning of these two terms do not apply to the present invention.
In a first aspect, the invention provides an antibody library for in-vitro identification of a specific antibody which binds a tumor cell, wherein each antibody in said library comprises
(i) a first antibody domain which binds a Tn- and/or STn-carbohydrate epitope of a glycoprotein of said tumor cell, and
(ii) a second antibody domain selected from a repertoire of second antibody domains, wherein the repertoire of second antibody domains comprises one or more second antibody domains which binds a peptide epitope of said glycoprotein of said tumor cell, wherein said specific antibody is specific for a combination of said carbohydrate epitope and said peptide epitope of said glycoprotein.
In one preferred embodiment, the first antibody domain is a VH-domain and the second antibody domain is a VL-domain. Hence, In a preferred embodiment, the invention provides an antibody library, wherein each antibody in the library comprises
(i) a VH-domain which binds a Tn- and/or STn-carbohydrate epitope of a glycoprotein of said tumor cell, and
(ii) a VL-domain selected from a repertoire of VL-domains, wherein the repertoire of VL-domains comprises one or more VL-domains which binds a peptide epitope of said glycoprotein of said tumor cell, for in-vitro identification of a specific antibody from said library which binds a tumor cell; wherein the specific antibody is specific for a combination of said carbohydrate epitope and said peptide epitope of said glycoprotein.
In a second aspect, the invention provides a nucleic acid library encoding the antibody library of the first aspect of the invention.
In a third aspect, the invention provides a method for identifying an antibody for targeting a tumor cell, comprising the steps of i) preparing an antibody library according to the first aspect of the invention, and ii) screening said library to identify one or more tumor targeting antibodies, preferably one or more specific tumor targeting antibodies. In a fourth aspect, the invention provides a method for identifying a glycopeptide target, said target comprising a Tn and/or STn epitope and a peptide epitope, such as a glycoprotein target of a cancer cell, said method comprising the steps of i) preparing an antibody library according to the first aspect of the invention, and ii) incubating the antibody library with a sample comprising the glycopeptide target, iii) analyzing one or more antibody-peptide complexes obtained from step ii) to identify the amino acid sequence of the peptide epitope of said glycopeptide target.
In a sixth aspect, the invention provides a specific tumor cell binding antibody, comprising
(i) a VH-domain which binds a Tn- and/or STn-carbohydrate epitope of a glycoprotein of the tumor cell, and
(ii) a VL-domain which binds a peptide epitope of said glycoprotein of said tumor cell, wherein the antibody is specific for a combination of said carbohydrate epitope and said peptide epitope of said glycoprotein, with the provisio that the antibody is not 5E5, 5F7 or 2D9.
Preferably, in one embodiment, the antibodies may be used in method of treatment and/or prevention of cancer.
In another embodiment, the antibodies may be used in diagnosing cancerous states in vivo or ex vivo in cell samples.
For this purpose, after being identified, isolated and/or generated from a library as described herein and/or using one of the methods described herein, the antibodies of the invention as described herein may also be subjected to one or more techniques known per se for improving one or more desired properties of the antibodies, such as (improved) affinity, (improved) potency or (reduced) immunogeneticy, and such improved antibodies form further aspects of the invention. For example and without limitation, in order to increase affinity and/or potency, the antibodies may be subjected to techniques for affinity maturation known per se. Also, potential immunogenicity may be reduced or removed by techniques for humanization known per se and/or using techniques for identifying potential immunogenic epitopes and then removing said potential epitopes by means of one or more suitable amino acid mutations, which again can be performed in a manner known per se. The amino acid of the antibody (or the sequence encoding such antibody) may for example also be subjected to techniques known per se for providing improved and/or increased expression in a desired host cell or host organism to be used for expression and/or production, which again will be clear to the skilled person, and may include known techniques for codon optimization. As will be clear to the skilled person, each of the preceding techniques may require or include a limited degree of trial-and-error which will be well within the skill of the artisan.
Preferably, but without being limited to a specific size beyond two different antibodles/clones, a library as described herein will usually comprise at least 10 different antibodies/clones or more, such as at least 50 different antibodies/clones or more, for example at least 100 different antibodies/clones or more. As will be clear to the skilled person, the upper limit for the size of a library as described herein is not critical and may be determined more by considerations such as desired diversity and practical considerations such as the size of library that can be easy and/or convient to generate, handle and screen. For example and without limitation, it is envisaged that a library as described herein may comprise more than 104 different clones, for example more than 105 different clones, such as more than 107 different clones and even up to 108, 109, IO10 or more different clones. Preferably, as mentioned herein, and in order to provide an advantageous degree of diversity, a library as used herein will comprise at least 1000 (103), such as at least 10.000 (104), preferably at least 100.000 (105), such as 105 or more, different sequences (i.e. antibodies with different VH/VL pairs).
Techniques for generating/constructing libraries of a size that are suitable for the purposes of the invention will be clear to the skilled person also based on the disclosure herein and for example include the techniques used in the following references: Ponsel et al., Molecules. 2011; 16(5): 3675-3700; Frenzel et al. (2014), Construction of human antibody gene libraries and selection of antibodies by phage display, in : Human monoclonal antibodies: methods and protocols: 215-243; Hutchings et al., (2001) Generation of naive human antibody libraries. In: Antibody engineering. Springer, pp 93-108; the review by Shim, BMB Rep. 2015; 48(9): 489-494; Mandrup et al., PLoS One, 8, (2013); Bai et al., PLoS ONE 2015, 10, e0141045; Hoet et al., Nat. Biotechnol 2005, 23, 344-348; Knappik et al., J. Mol. Biol. 2000, 296, 57-86; Kugler et al., BMC Biotechnol. 2015, 15, 10; Prassler et al., J. Mol. Biol. 2011, 413, 261-278; Soderlind et al., Nat. Biotechnol. 2000, 18, 852-856; Tiller et al., Int. J. Mol. Sci. 2022, 23, 6255; and Valadon et al., mAbs 2019, 11, 516-531. Other suitable techniques will be clear to the skilled person.
As will also be clear to the skilled person based on the disclosure herein, generating such library will generally comprise combining a VH sequence that has been chosen for its ability to specifically bind to a Tn or STn epitope (such as one of the VH sequences described herein or referred to herein) with a repertoire of VL sequences that provides the desired size and diversity to the library. For example, as further described herein, such a VL repertoire may be a collection of naive (for example obtained from a naive library generated from mouse B-cells or human B-cells), pre-immune, synthetic or semisynthetic sequences.
Also, such a library may be in any suitable format, including DNA, RNA or protein, and may be in the form of a library in which the VH and VL sequences are present in a suitable format/vector that allows for suitable expression or display of the antibodies. This may for example be a phage library or yeast library, depending on the technique(s) that are intended to be used for screening the library. Suitable screening techniques (and suitable library formats for use in such screening techniques) will be clear to the skilled person, and for example and without limitation include (techniques and libraries for) phage display, ribosome display, yeast display or display using suitable mammalian cell systems.
It should also be noted that, for the numbering of the amino acid residues in a VH and VL domain as described herein, and also for defining the CDRs in such a VH and VL domains, generally the numbering scheme of Kabat (Kabat et aL, Sequences of Immunoglobulin Chains: Tabulation and Analysis of Amino Acid Sequences of Precursors, V-regions, C-regions, J-Chain and BP-Microglobulins, 1979. Department of Health, Education, and Welfare, Public Health Service, National Institutes of Health (1979)) will be used, unless indicated otherwise. However, it will be clear to the skilled person that other schemes exist (for example, the Chotia, IMGT and AbM schemes) and the skilled person will be able to apply these schemes to the VH and VL sequences described herein.
Also, as used herein, the terms "CDR", "CDR1", "CDR2" and "CDR3" have their usual meaning in the art and are, unless specifically, as defined according to Kabat.
Furthermore, when comparing two amino acid sequences, the term "amino acid difference" as used herein refers to an insertion, deletion or substitution of a single amino acid residue on a position of the first sequence, compared to the second sequence; it being understood that two amino acid sequences can contain one, two or more such amino acid differences;
In the context of the present invention and claims, when referring to the ability of an antibody to bind to an antigen, such binding is preferably specific binding, which typically means that such an antibody will bind to its antigen with a dissociation constant (KD) of 10-5 to 10 12 moles/liter or less, and preferably IO-7 to 10 12 moles/liter or less and more preferably 10-8 to 10 12 moles/liter (i.e. with an association constant (KA) of 105 to 1012 liter/ moles or more, and preferably 107 to 1012 liter/moles or more and more preferably 108 to 1012 liter/moles). Any KD value greater than 104 mol/liter (or any KA value lower than 104 M-l) liters/mol is generally considered to indicate non-specific binding. Preferably, an antibody of the invention will bind to the desired antigen with an affinity less than 500 nM, preferably less than 200 nM, more preferably less than 10 nM, such as less than 500 pM. Specific binding of an antigen-binding protein to an antigen or antigenic determinant can be determined in any suitable manner known per se, including, for example, Scatchard analysis and/or competitive binding assays, such as radioimmunoassays (RIA), enzyme immunoassays (EIA) and sandwich competition assays, and the different variants thereof known per se in the art; as well as the other techniques mentioned herein.
It should also be noted that, generally and as will be clear to the skilled person, the VH domains and VL domains referred to herein will be part of, and in some cases in essence may have to be part of, a VH/VL pair in order to form a complete binding site. It will also be clear to the skilled person that, for this reason, it may not in all cases be practicable or possible to determine whether an individual VH domain or VL domain can bind to an antigen or epitope when it is not part of a VH/VL pair. Thus, when in the present specification and claims, a VH domain or VL domain is referred to as being able to (specifically) bind to an epitope, antigen or protein, this includes both a situation where such domain can bind (and/or such binding can be determined or measured) when the domain is in isolated form (i.e. not part of a VH/VL pair) or where binding of such a domain takes place (and/or can be determined or measured) when such a domain is part of a suitable VH/VL pair as described herein (i.e. a VH/VL pair as is present in the antibodies described herein).
In a further aspect, the invention provides a method for generating antibodies against a target that is present on the surface of a cancer cell (and/or is expressed by a cancer cell on its surface), in which said target is present on/expressed on the cancer cell in a glycosylated form (and in particular in a glycosylated form that contains, comprises or presents one or more Tn ot STn epitopes, as as described herein), which method at least comprises the following steps:
Constructing or providing a library of antibodies (or a library of sequences that encode antibodies), in which each antibody in the library comprises a VH domain and VL domain, in which each such VH domain is (chosen to be) capable of binding (and preferably specifically binding, as further defined herein) to a Tn epitope or a STn epitope, and in which the library is preferably such that it comprises (a collection or repertoire of) at least 1000 (103), such as at least 10.000 (104), preferably at least 100.000 (IO5), such as IO5 or more, different VL domains such that, upon each of said VL domains combining with said VH domain as part of an antibody that can be obtained from said library, said library comprises/provides (a collection or repertoire of) at least 1000 (103), such as at least 10.000 (104), preferably at least 100.000 (105), such as 105 or more, different antibodies (i.e. antibodies having different VH/VL combinations); screening said library against said target, and preferably screening said library against said target in which said target is in a glycosylated form, and even more preferably screening said library against said target in which said target is in a glycosylated form that contains, comprises or presents one or more Tn or STn epitopes; identifying, generating and/or isolating one or more antibodies (or sequences encoding such antibodies) that are capable of specifically binding to said target, and that are preferably capable of specifically binding to said target in a glycosylated form, and more preferably capable of specifically binding to said target in a glycosylated form that contains, comprises or presents one or more Tn or STn epitopes; and optionally applying one or more steps for improving one or desired properties of the antibody/antibodies thus obtained, for example by means of humanization, affinity maturation, removing potential immunogenic epitopes and/or optimizing the sequence for expression or production in a desired host cell or host organism.
Based on the disclosure herein, it will be clear to the skilled person that some of the libraries of the invention may comprise a single VH sequence (i.e. chosen to specifically bind to a Tn epitope or STn epitope, as further described herein) that is combined with a suitable collection or repertoire of different VL sequences, for example at least 1000 (103), such as at least 10.000 (104), preferably at least 100.000 (105), such as 105 or more different VL sequences. Other libraries may contain two or more different VH sequences, in which each such VH sequence has been chosen to specifically bind to a Tn epitope or to an STn epitope (in which said VH sequences are again suitably combined with a suitable collection or repertoire of VL sequences, as further described herein).
Suitable VH sequences that can be used to generate the libraries of the invention described herein will be clear to the skilled person based on the disclosure herein. For example, suitable candidates for VH sequences that may be useful in constructing the libraries of the invention include, but are not limited to, VH sequences that are present in and/or have been derived from antibodies that have been raised against glycosylated proteins that are present and/or expressed on the surface of cancer cells, in particular where such proteins are known to contain, comprise or present Tn or STn epitopes. Such VH sequences may then be tested for their ability to be used as a VH sequence in constructing the libraries described herein. For example and without limitation, in the invention, the VH sequences that are present in the antibodies 5E5, anti-CD43, 2D9, 5F7, G2D11, 3F1, 83D4, 15G9, 1E3, MLS128, 16E12.1D9.1B11 and/or 1A5-2C9 (all referred to supra) may be tested for their ability to serve as a VH sequence in the libraries of the invention (and, by extension, in antibodies generated from such a library); and such VH sequences (or VH sequences comprising the CDRs that are present in these antibodies) may then be used as VH sequences in constructing the libraries provided by the invention.
The libraries of the invention are preferably such that all or essentially all VH sequences that are present in the library (and/or that have been used in constructing the library) are capable of binding to either a Tn or STn epitope. Furthermore, the library of the invention is also preferably such that at least 90%, such as at least 95%, and more preferably all or essentially all antibodies that are obtained by means of the methods described herein (i.e. from screening the library) contain a VH sequence that is capable of binding to a Tn or STn epitope; and even more preferably contain a VH sequence that confers, to the antibody or antibodies obtained from the library, the ability to specifically bind to (the TN or STn part) of an epitope or antigen that comprises a Tn or STn epitope.
The invention also provides some VH sequences that have been found to be particularly suited for use as VH sequences in the libraries provided by the invention (and again, by extension, in antibodies that can be generated using such libraries). These are the VH sequence of SEQ ID NO: 1 (which can be used to generate libraries that have specificity for Tn epitopes/proteins comprising Tn epitopes) and the VH sequence of SEQ ID NO: 28 (which can be used to generate libraries that have specificity for STn epitopes/proteins comprising STn epitopes). Said VH sequences (and other suitable VH sequences having the same CDRs as the VH sequences of SEQ ID NO: 1 or SEQ ID NO: 28, respectively), libraries of the invention comprising and/or based on such VH sequences, methods for generating such librariesm and antibodies identified using, generated using and/or isolated from such libraries form further preferred aspects of the invention.
Further suitable VH sequences will be clear to the skilled person based on the disclosure herein and include: (i) amino acid sequences having at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 1, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 1; and preferably such amino acid sequences that comprise the amino acid residues H32, A33, H35, Y50, and S99, with respect to SEQ ID NO. 1 (which amino acid sequences will, as mentioned herein, provide recognition support for mono-Tn epitopes); and (ii) amino acid sequences having at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 28, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 28; and preferably such amino acid sequences that comprise the amino acid residues T28, T30, H32, A33, H35, Y50, S99, L101, A102 and L103, with respect to SEQ ID NO. 28 (which amino acid sequences will, as mentioned herein, provide recognition support for mono- STn epitopes). Again, such VH sequences, libraries of the invention comprising and/or based on such VH sequences, methods for generating such libraries and antibodies identified using, generated using and/or isolated from such libraries form further preferred aspects of the invention.
Thus, in a further aspect, the invention provides a method for generating antibodies against a target that is present on the surface of a cancer cell (and/or is expressed by a cancer cell on its surface), in which said target is present on/expressed on the cancer cell in a glycosylated form (and in particular in a glycosylated form that contains, comprises or presents one or more Tn ot STn epitopes, as as described herein), which method at least comprises the following steps:
Constructing or providing a library of antibodies (or a library of sequences that encode antibodies ), in which each antibody in the library comprises a VH domain and VL domain, in which each such VH domain is (i) an amino acid sequence having at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 1, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 1; and preferably such an amino acid sequence that comprises the amino acid residues H32, A33, H35, Y50, and S99, with respect to SEQ ID NO. 1; and/or (ii) an amino acid sequence having at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 28, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 28; and preferably such an amino acid sequences that comprises the amino acid residues T28, T30, H32, A33, H35, Y50, S99, L101, A102 and L103, with respect to SEQ ID NO. 28; and in which the library is preferably such that it comprises (a collection or repertoire of) at least 1000 (103), such as at least 10.000 (104), preferably at least 100.000 (105), such as 105 or more, different VL domains such that, upon each of said VL domains combining with said VH domain as part of an antibody that can be obtained from said library, said library comprises/provides (a collection or repertoire of) at least 1000 (103), such as at least 10.000 (104), preferably at least 100.000 (105), such as 105 or more, different antibodies (i.e. antibodies having different VH/VL combinations); screening said library against said target, and preferably screening said library against said target in which said target is in a glycosylated form, and even more preferably screening said library against said target in which said target is in a glycosylated form that contains, comprises or presents one or more Tn or STn epitopes; identifying, generating and/or isolating one or more antibodies (or sequences encoding such antibodies) that are capable of specifically binding to said target, and that are preferably capable of specifically binding to said target in a glycosylated form, and more preferably capable of specifically binding to said target in a glycosylated form that contains, comprises or presents one or more Tn or STn epitopes; and optionally applying one or more steps for improving one or desired properties of the antibody/antibodies thus obtained, for example by means of humanization, affinity maturation, removing potential immunogenic epitopes and/or optimizing the sequence for expression or production in a desired host cell or host organism.
In yet another aspect, the invention provides a method for generating antibodies against a target that is present on the surface of a cancer cell (and/or is expressed by a cancer cell on its surface), in which said target is present on/expressed on the cancer cell in a glycosylated form (and in particular in a glycosylated form that contains, comprises or presents one or more Tn ot STn epitopes, as as described herein), which method at least comprises the following steps:
Constructing or providing a library of antibodies (or a library of sequences that encode antibodies), in which each antibody in the library comprises a VH domain and VL domain, in which each such VH domain has the amino acid sequence of SEQ ID NO. 1 and/or the amino acid sequence of SEQ ID NO. 28; and in which the library is preferably such that it comprises (a collection or repertoire of) at least 1000 (103), such as at least 10.000 (104), preferably at least 100.000 (105), such as 105 or more, different VL domains such that, upon each of said VL domains combining with said VH domain as part of an antibody that can be obtained from said library, said library comprises/provides (a collection or repertoire of) at least 1000 (103), such as at least 10.000 (104), preferably at least 100.000 (IO5), such as IO5 or more, different antibodies (i.e. antibodies having different VH/VL combinations); screening said library against said target, and preferably screening said library against said target in which said target is in a glycosylated form, and even more preferably screening said library against said target in which said target is in a glycosylated form that contains, comprises or presents one or more Tn or STn epitopes; identifying, generating and/or isolating one or more antibodies (or sequences encoding such antibodies) that are capable of specifically binding to said target, and that are preferably capable of specifically binding to said target in a glycosylated form, and more preferably capable of specifically binding to said target in a glycosylated form that contains, comprises or presents one or more Tn or STn epitopes; and optionally applying one or more steps for improving one or desired properties of the antibody/antibodies thus obtained, for example by means of humanization, affinity maturation, removing potential immunogenic epitopes and/or optimizing the sequence for expression or production in a desired host cell or host organism.
In yet another aspect, the invention provides a method for generating antibodies against a target that is present on the surface of a cancer cell (and/or is expressed by a cancer cell on its surface), in which said target is present on/expressed on the cancer cell in a glycosylated form (and in particular in a glycosylated form that contains, comprises or presents one or more Tn ot STn epitopes, as as described herein), which method at least comprises the following steps:
Constructing or providing a library of antibodies (or a library of sequences that encode antibodies), in which each antibody in the library comprises a VH domain and VL domain, in which each such VH domain comprises: (i) a CDR1 having the amino acid sequence DHAIH (SEQ ID NO: 155) or an amino acid sequence that has a single amino acid difference (as defined herein) with the amino acid sequence DHAIH (SEQ ID NO: 155); and (ii) a CDR2 having the amino acid sequence YISPGNDDIKYNEKFKG (SEQ ID NO: 156) or an amino acid sequence that has 3, 2 or 1 (and preferably 2 or 1, and most preferably 1) amino acid differences (as defined herein) with the amino acid sequence YISPGNDDIKYNEKFKG (SEQ ID NO: 156); and (iii) a CDR3 having the amino acid sequence SLPGTFDY (SEQ ID NO: 157) or an amino acid sequence that has a single amino acid difference (as defined herein) with the amino acid sequence SLPGTFDY (SEQ ID NO: 157); and/or in which each such VH domain comprises: (i) a CDR1 having the amino acid sequence DHAIH (SEQ ID NO: 155) or an amino acid sequence that has a single amino acid difference (as defined herein) with the amino acid sequence DHAIH (SEQ ID NO: 155); and (ii) a CDR2 having the amino acid sequence YISPGNDDIKYNEKFKG (SEQ ID NO: 156) or an amino acid sequence that has 3, 2 or 1 (and preferably 2 or 1, and most preferably 1) amino acid differences (as defined herein) with the amino acid sequence YISPGNDDIKYNEKFKG (SEQ ID NO: 156); and (iii) a CDR3 having the amino acid sequence SLLALDY (SEQ ID NO: 158) or an amino acid sequence that has a single amino acid difference (as defined herein) with the amino acid sequence SLLALDY (SEQ ID NO: 158); and in which the library is preferably such that it comprises (a collection or repertoire of) at least 1000 (103), such as at least 10.000 (104), preferably at least 100.000 (IO5), such as IO5 or more, different VL domains such that, upon each of said VL domains combining with said VH domain as part of an antibody that can be obtained from said library, said library comprises/provides (a collection or repertoire of) at least 1000 (103), such as at least 10.000 (104), preferably at least 100.000 (105), such as IO5 or more, different antibodies (i.e. antibodies having different VH/VL combinations); screening said library against said target, and preferably screening said library against said target in which said target is in a glycosylated form, and even more preferably screening said library against said target in which said target is in a glycosylated form that contains, comprises or presents one or more Tn or STn epitopes; identifying, generating and/or isolating one or more antibodies (or sequences encoding such antibodies) that are capable of specifically binding to said target, and that are preferably capable of specifically binding to said target in a glycosylated form, and more preferably capable of specifically binding to said target in a glycosylated form that contains, comprises or presents one or more Tn or STn epitopes; and optionally applying one or more steps for improving one or desired properties of the antibody/antibodies thus obtained, for example by means of humanization, affinity maturation, removing potential immunogenic epitopes and/or optimizing the sequence for expression or production in a desired host cell or host organism.
As further described herein, in the libraries of the invention: when a VH sequence comprises a CDR1 that is an amino acid sequence that has a single amino acid difference (as defined herein) with the amino acid sequence DHAIH (SEQ ID NO: 155), then such a CDR1 preferably comprises the amino acid residues H32, A33, and H35; when a VH sequence comprises a CDR2 that is an amino acid sequence that has 3, 2 or 1 (and preferably 2 or 1, and most preferably 1) amino acid differences (as defined herein) with the amino acid sequence YISPGNDDIKYNEKFKG (SEQ ID NO: 156), then such a CDR2 preferably comprises the amino acid residues Y5O, S52, N55, and D57; and when a VH sequence comprises a CDR3 that is an amino acid sequence that has a single amino acid difference (as defined herein) with the amino acid sequence SLPGTFDY (SEQ ID NO: 157), then such a CDR3 preferably comprises the amino acid residue S99; or when a VH sequence comprises a CDR3 that is an amino acid sequence that has a single amino acid difference (as defined herein) with the amino acid sequence SLLALDY (SEQ ID NO: 158), then such a CDR3 preferably comprises the amino acid residues S99, L1O1, A102 and L1O3.
Based on the disclosure herein, it will be clear to the skilled person that, in the practice of the invention, it is possible to use a library that both comprises one or more VH sequences that have been chosen to confer specificity for a Tn epitope (or a protein that contains, comprises or presents a Tn epitope) and also comprises one or more VH sequences that have been chosen to confer specificity for an STn epitope (or a protein that contains, comprises or presents a STn epitope). For example, such a library can be suitably constructed using (or suitably comprise) both the VH sequence of SEQ ID NO: 1 (or a VH sequence based on the VH sequence of SEQ ID NO: 1, as further described herein) and the VH sequence of SEQ ID NO: 28 (or a VH sequence based on the VH sequence of SEQ ID NO: 28, as further described herein). Such a library can for example be used to generate and/or screen for antibodies that are expressed on a cancer cell, irrespective of whether said protein comprises or presents a Tn epitope, an STn epitope, or both. Thus, one aspect of the invention relates to a library as described herein that contains at least one VH sequence that confers, to the antibodies that can be obtained from such a library, specificity for a Tn epitope and further contains at least one VH sequence that confers, to the antibodies that can be obtained from such a library, specificity for a STn epitope. Again, as further described herein, in such a library, the VH domain conferring specificity to the Tn epitope may be the VH domain of SEQ ID NO: 1 (or a variant thereof as described herein and/or a VH domain having CDRs that are based on the CDRs of SEQ ID NO: 1, as also described herein) and the VH domain conferring specificity to the STn epitope may be the VH domain of SEQ ID NO: 28 (or a variant thereof as described herein and/or a VH domain having CDRs that are based on the CDRs of SEQ ID NO:28, as also described herein)
It will also be clear to the skilled person that, when the target contains or is expected to contain one or more Tn epitopes, and/or when it is desired to generate/obtain an antibodies that are capable of (specifically) binding to a target when said target presents a Tn epitope, the methods of the invention will usually involve the use of a library that only contains one or more VH sequence(s) that are capable of specifically binding to a Tn epitope (and/or have been chosen based on their ability to specifical bind to a Tn epitope). Again, as further described herein, in such a library, the VH domain conferring specificity to the Tn epitope may be the VH domain of SEQ ID NO: 1 (or a variant thereof as described herein and/or a VH domain having CDRs that are based on the CDRs of SEQ ID NO: 1, as also described herein); and libraries that only contain VH sequences that are capable of specifically binding to a Tn epitope (and/or have been chosen based on their ability to specifical bind to a Tn epitope) form further aspects of the invention.
Similarly, when the target contains or is expected to contain one or more STn epitopes, and/or when it is desired to generate/obtain an antibodies that are capable of (specifically) binding to a target when said target presents a STn epitope, the methods of the invention will usually involve the use of a library that only contains one or more VH sequence(s) that are capable of specifically binding to a STn epitope (and/or have been chosen based on their ability to specifical bind to a STn epitope). Again, as further described herein, in such a library, the VH domain conferring specificity to the STn epitope may be the VH domain of SEQ ID NO: 28 (or a variant thereof as described herein and/or a VH domain having CDRs that are based on the CDRs of SEQ ID NO:28, as also described herein); and libraries that only contain VH sequences that are capable of specifically binding to a STn epitope (and/or have been chosen based on their ability to specifical bind to a STn epitope) form further aspects of the invention.
In each of the libraries of the invention as described herein, the (collection or repertoire of different) VL sequences that can be present/included in the library may be provided and/or have been generated in any suitable manner known per se, and may for example be a collection or repertoire of naive VL sequences (for example derived from a naive library of antibody sequences obtained from mouse B-cells or human B-cells), pre- immume VL sequences, synthetic VL sequences and/or semi-synthetic VL sequences; or any combination of the foregoing. Methods and techniques for providing (sequences encoding) such a collection or repertoire of VL sequences, and for suitably including them in a library as described herein, will be clear to the skilled person (reference is for example made to following references: Ponsel et al.; Frenzel et al.; Hutchings et al., (2001); Shim; Mandrup et al.; Bai et al.; Hoet et al.; Knappik et al.; Kugler et al.; Prassler et al.; Soderlind et al.; Tiller et al.; and Valadon et al.; all supra). In the practice of the invention, these and other suitable techniques may be used or suitably adapted to provide a library in which the desired VL sequences are suitably combined with one or more VH sequences that have been chosen to specifically bind to a Tn or STn epitope, respectively; so as to provide a library of the invention as further described herein, in which the size and/or diversity of the library is mainly provided by (the collection or repertoire of) the different VL sequences that are present in the library.
Thus, the libraries of the invention may contain or comprise a collection or repertoire of naive VL sequences, a collection or repertoire of synthetic VL sequences, or a collection or repertoire of semi-synthetic VL sequences (with libraries based on a collection or repertoire of naive VL sequences forming one preferred aspect of the invention), in each case combined with a VH sequence that is as further described herein. In this respect, it should be noted that, while the presence/use of an immune repertoire of VL sequences is not excluded from the invention in its broadest sense, the use of an immune repertoire will usually be less preferred, as the VL sequences derived from an immune VH/VL repertoire may often requiring pairing with the specific VH sequence that they are associated with in the immune repertoire in order to provide the degree of specificity that is intended for the purposes of the present invention.
Also, as further described herein, in the antibodies that can be obtained from each of these libraries (i.e. by means of screening and selection as further described herein), when the VH sequence that is present in such antibody confers specificity for a Tn epitope, the VL domain in the antibody is preferably such that it is capable of binding to (part of) the peptide backbone that, in the intended or desired target, is associated with the Tn epitope to which VH sequence in the antibody can bind. Similarly, and again as further described herein, in the antibodies that can be obtained from each of these libraries (i.e. by means of screening and selection as further described herein), when the VH sequence that is present in such antibody confers specificity for a STn epitope, the VL domain in the antibody is preferably such that it is capable of binding to (part of) the peptide backbone that, in the intended or desired target, is associated with the STn epitope to which VH sequence in the antibody can bind. In each such case, (the sequence of) the VH domain is preferably such that it essentially does not contribute to the binding of the antibody to the peptide epitope (i.e. to the associated part of the peptide backbone of the target).
As further described herein, the methods and libraries provided by the invention can in particular be used to obtain, identify and/or generate (sequences encoding) antibodies against proteins/targets that are present on/expressed on cancer cells, in particular against proteins/targets that are present on/expressed on cancer cells in a glycosylated form, and more in particular against proteins/targets that are present on/expressed on cancer cells in a glycosylated form where said glycosylated form comprises, contains or presents one or more Tn and/or STn epitopes (as further described herein).
Such proteins/targets, as well as the types of cancer cells on which they are expressed (and sometimes overexpressed compared to healthy cells and/or other cancer cells) and the cancer types with which such proteins, targets and cells are associated will be clear to the skilled person, and for example and without limitation include: EGFR, VEGFR, HER2, CD37, FLT3, FGFR, CD19, CD22, CD27, CD25, CD30, CD33, CD38, CD43, Mesothelin, PD-L1, CD44, Podocalyxin (TRA1.60/80), CD133, CD90, CD326, Cripto-1, ABCG2, CD24, CD49, Notch2, CD146 (MUC18), CD10, CD117, CD26, CXCR4, CD34, CD271, CD13, CD56, CD105, LGR5, CD114, CD54, CXCR1, TIM-3, CD55, DLL-4, CD96, CD29, CD9, CD166, CD44, ABCB5, Notch3, CD123, MUC1, MUC4, MUC 13, MUC16, MUC17, MUC21, CD33, CD40, Claudin-18, Integrin alpha-3, CK7, CK20, CXCR2, CXCR4, Integrin alpha-5, CA9, EPCAM, MET, MMP14, DDR1, NRP1, LAMP1-4, CD99, ALCAM, SDC1, 2, 3, 4, Syndecans (1-4), ITA5, ROBO1, Nectin-4, Nectin-2, LRP1, CD70, Podoplanin, IGF-1R, ABCG2, IBT1, ALK, TEM1, HVEM, TERT, LYPD3, GPR56, IL6RA, DDR1, LIFR, GPR64, S1PR1-4, LRCHD2, NMB, Plexinl/2, PTHR, Semaphorins, Slit3, TACD2, VGFC, VLDLR, VTN, LAG-3, Fnl4 and S1PR3.
Generally, when it is desired to generate an antibody for treating a specific type of cancer, a person skilled in the field of oncology will be able to select a protein or target that will be associated with cancer cells that are involved in said type of cancer and then use antibodies against said target(s) that have been generated using the methods and libraries described herein for treating said type of cancer. Alternativelty, the skilled person will be able to determine which (glycosylated) proteins are overexpressed on the (type of) cancer cell involved and then use the methods and libraries of the invention to generate one or more antibodies against saud proteins. It is also envisaged that the latter aspect of the invention may find use in the area of so-called "personalized medicine", by allowing the skilled person to generate antibodies that are directed against proteins that are expressed on cancer cells that have been obtained from the patient to be treated, and then use said antibodies to treat said patient.
In one aspect of the invention, the methods and libraries of the invention are used to generate antibodies against a protein/target belonging to the MUC family that are expressed (in a glycosylated form) on a cancer cell. The invention also relates to methods for treating cancer which comprise administering, to a patient in need thereof, one or more a therapeutically effective amounts of an antibody that can be obtained or have been obtained from the libraries and/or using the methods described herein. In a specific aspect, the invention further relates to methods for treating cancer which comprise administering, to a patient suffering from cancer, one or more a therapeutically effective amounts of an antibody that can be obtained or have been obtained from the libraries and/or using the methods described herein, in which said antibody is directed against a (glycosylated) protein that is present on/expressed by a cancer cell that is present in the body of the patient to be treated. The invention also relates antibodies that can be obtained or have been obtained from the libraries and/or using the methods described herein for the treatment of cancer. In each asopect, the antibodies used can be as further described herein. In one specific aspect, such antibodies are directed against a protein/target belonging to the MUC family (as further described herein).
With respect to the generation, selection and/or use of a specific antibody as described herein for the treatment of a particular type of cancer, it should also be noted that, as for example described in Romer, T. B. et al., Brit J Cane 125, 1239 1250 (2021); Jiang, Y. et al., J Cell Mol Med 22, 4875 4885 (2018); and
Tsuchiya et al. Breast Cane 6, 175-180 (1999), that in certain types of cancer, the proteins expressed on the surface of the cancer cell mainly or predominantly present Tn antigens, whereas the cells of other types of cancer mainly or predominantly express proteins that present STn antigens. For example, as mentioned by Romer et al., Tn antigen expression is mainly observed in breast, colorectal and pancreatic tumors while STn antigen was highly expressed in colorectal and pancreatic tumors.
Accordingly, a further aspect of the invention relates to the use of antibody that can be obtained or have been obtained from the libraries and/or using the methods described herein for the treatment of cancer, where the cells of the tumor involved and/or the type of cancer involved mainly express proteins on their surface that contain or present Tn antigens, in which the antibody used comprises a VH domain that confers specificity for a Tn antigen and/or wherein the antibody has been obtained from a library as described herein that was constructed using VH sequences that confer specificity for a Tn antigen. As will be clear to the skilled person based on the disclosure herein, the VL sequence present in such an antibody will be as further described herein and may in particular confer specificity for (the protein backbone of) a protein present or expressed on the cancer cell that contains or presents the Tn antigen. Another aspect of the invention relates to the use of antibody that can be obtained or have been obtained from the libraries and/or using the methods described herein for the treatment of cancer, where the cells of the tumor involved and/or the type of cancer involved mainly express proteins on their surface that contain or present STn antigens, in which the antibody used comprises a VH domain that confers specificity for a STn antigen and/or wherein the antibody has been obtained from a library as described herein that was constructed using VH sequences that confer specificity for a STn antigen. As will be clear to the skilled person based on the disclosure herein, the VL sequence present in such an antibody will be as further described herein and may in particular confer specificity for (the protein backbone of) a protein present or expressed on the cancer cell that contains or presents the STn antigen.
In further aspects, the invention relates to antibodies that can be obtained or have been obtained from the libraries described herein and/or using the methods described herein.
In one specific aspect, such antibodies contain, as a VH domain, an amino acid sequence that is the VH sequence of SEQ ID NO: 1, or a variant of the VH sequence of SEQ ID NO: 1 (which variant is preferably as further described herein) or a VH sequence that has CDRs that are the same as the CDR sequences that are present in the VH sequence of SEQ ID NO: 1 or that have been derived from the CDR sequences that are present in the VH sequence of SEQ ID NO: 1. Said CDR sequences present in the VH sequence of SEQ ID NO: 1 are as follows:
CDR1 = DHAIH (SEQ ID NO: 155);
CDR2 = YISPGNDDIKYNEKFKG (SEQ ID NO: 156);
CDR3 = SLPGTFDY (SEQ ID NO: 157).
In another specific aspect, such antibodies contain, as a VH domain, an amino acid sequence that is the VH sequence of SEQ ID NO: 28, or a variant of the VH sequence of SEQ ID NO: 28 (which variant is preferably as further described herein) or a VH sequence that has CDRs that are the same as the CDR sequences that are present in the VH sequence of SEQ ID NO: 28 or that have been derived from the CDR sequences that are present in the VH sequence of SEQ ID NO: 28. Said CDR sequences present in the VH sequence of SEQ ID NO: 28 are as follows:
CDR1 = DHAIH (SEQ ID NO: 155);
CDR2 = YISPGNDDIKYNEKFKG (SEQ ID NO: 156);
CDR3 = SLLALDY (SEQ ID NO: 158). Thus, in a further aspect, the invention relates to an antibody, and in particular an antibody that can be obtained or has been obtained using the methods described herein and/or from a library as described herein, which antibody comprises a VH domain and a VL domain, in which the VH domain comprises: a CDR1 having the amino acid sequence DHAIH (SEQ ID NO: 155) or an amino acid sequence that has a single amino acid difference (as defined herein) with the amino acid sequence DHAIH (SEQ ID NO: 155); and
- a CDR2 having the amino acid sequence YISPGNDDIKYNEKFKG (SEQ ID NO: 156) or an amino acid sequence that has 3, 2 or 1 (and preferably 2 or 1, and most preferably 1) amino acid differences (as defined herein) with the amino acid sequence YISPGNDDIKYNEKFKG (SEQ ID NO: 156); and a CDR3 having the amino acid sequence SLPGTFDY (SEQ ID NO: 157) or an amino acid sequence that has a single amino acid difference (as defined herein) with the amino acid sequence SLPGTFDY (SEQ ID NO: 157); and in which the VL domain is as further described herein, and is in particular such that said VL domain is capable of binding to the peptide backbone of a glycosylated protein that is present or expressed on the surface of a cancer cell (and more in particular capable of binding to the peptide backbone of a glycosylated protein that is present or expressed on the surface of a cancer cell, where said glycosylated protein contains or presents a Tn antigen) and/or such that said VL domain confers upon the antibody the ability to specifically bind to a protein that is present or expressed on the surface of a cancer cell (where, in particular, said protein is glycosylated and more in particular glycosylated such that it contains or presents a Tn antigen).
As will be clear to the skilled person based on the disclosure herein, in the antibodies according to the preceding aspect: when the VH sequence that is present in the antibody comprises a CDR1 that is an amino acid sequence that has a single amino acid difference (as defined herein) with the amino acid sequence DHAIH (SEQ ID NO: 155), then such a CDR1 preferably comprises the amino acid residues H32, A33, and H35; when the VH sequence that is present in the antibody a CDR2 that is an amino acid sequence that has 3, 2 or 1 (and preferably 2 or 1, and most preferably 1) amino acid differences (as defined herein) with the amino acid sequence YISPGNDDIKYNEKFKG (SEQ ID NO: 156), then such a CDR2 preferably comprises the amino acid residues Y50, S52, N55, and D57; and when the VH sequence that is present in the antibody comprises a CDR3 that is an amino acid sequence that has a single amino acid difference (as defined herein) with the amino acid sequence SLPGTFDY (SEQ ID NO: 157), then such a CDR3 preferably comprises the amino acid residue S99. In a specific but non-limiting aspect, the invention relates to an antibody, and in particular an antibody that can be obtained or has been obtained using the methods described herein and/or from a library as described herein, which antibody comprises a VH domain and a VL domain, in which the VH domain comprises a CDR1 having the amino acid sequence DHAIH and a CDR2 having the amino acid sequence YISPGNDDIKYNEKFKG and a CDR3 having the amino acid sequence SLPGTFDY (SEQ ID NO: 157); and in which the VL domain is as further described herein, and is in particular such that said VL domain is capable of binding to the peptide backbone of a glycosylated protein that is present or expressed on the surface of a cancer cell (and more in particular capable of binding to the peptide backbone of a glycosylated protein that is present or expressed on the surface of a cancer cell, where said glycosylated protein contains or presents a Tn antigen) and/or such that said VL domain confers upon the antibody the ability to specifically bind to a protein that is present or expressed on the surface of a cancer cell (where, in particular, said protein is glycosylated and more in particular glycosylated such that it contains or presents a Tn antigen).
In a further aspect, the invention relates to an antibody, and in particular an antibody that can be obtained or has been obtained using the methods described herein and/or from a library as described herein, which antibody comprises a VH domain and a VL domain, in which the VH domain comprises: a CDR1 having the amino acid sequence DHAIH (SEQ ID NO: 155) or an amino acid sequence that has a single amino acid difference (as defined herein) with the amino acid sequence DHAIH (SEQ ID NO: 155); and
- a CDR2 having the amino acid sequence YISPGNDDIKYNEKFKG (SEQ ID NO: 156) or an amino acid sequence that has 3, 2 or 1 (and preferably 2 or 1, and most preferably 1) amino acid differences (as defined herein) with the amino acid sequence YISPGNDDIKYNEKFKG (SEQ ID NO: 156); and a CDR3 having the amino acid sequence SLLALDY (SEQ ID NO: 158) or an amino acid sequence that has a single amino acid difference (as defined herein) with the amino acid sequence SLLALDY (SEQ ID NO: 158) and in which the VL domain is as further described herein, and is in particular such that said VL domain is capable of binding to the peptide backbone of a glycosylated protein that is present or expressed on the surface of a cancer cell (and more in particular capable of binding to the peptide backbone of a glycosylated protein that is present or expressed on the surface of a cancer cell, where said glycosylated protein contains or presents a STn antigen) and/or such that said VL domain confers upon the antibody the ability to specifically bind to a protein that is present or expressed on the surface of a cancer cell (where, in particular, said protein is glycosylated and more in particular glycosylated such that it contains or presents a STn antigen).
As will be clear to the skilled person based on the disclosure herein, in the antibodies according to the preceding aspect: when the VH sequence that is present in the antibody comprises a CDR1 that is an amino acid sequence that has a single amino acid difference (as defined herein) with the amino acid sequence DHAIH (SEQ ID NO: 155), then such a CDR1 preferably comprises the amino acid residues H32, A33, and H35; when the VH sequence that is present in the antibody a CDR2 that is an amino acid sequence that has 3, 2 or 1 (and preferably 2 or 1, and most preferably 1) amino acid differences (as defined herein) with the amino acid sequence YISPGNDDIKYNEKFKG (SEQ ID NO: 156), then such a CDR2 preferably comprises the amino acid residues Y50, S52, N55, and D57; and when the VH sequence that is present in the antibody comprises a CDR3 that is an amino acid sequence that has a single amino acid difference (as defined herein) with the amino acid sequence SLLALDY (SEQ ID NO: 158), then such a CDR3 preferably comprises the amino acid residues S99, L101, A102 and L103.
In a further aspect, the invention relates to an antibody, and in particular an antibody that can be obtained or has been obtained using the methods described herein and/or from a library as described herein, which antibody comprises a VH domain and a VL domain, in which the VH domain comprises a CDR1 having the amino acid sequence DHAIH (SEQ ID NO: 155) and a CDR2 having the amino acid sequence YISPGNDDIKYNEKFKG (SEQ ID NO: 156) and a CDR3 having the amino acid sequence SLLALDY (SEQ ID NO: 158); and in which the VL domain is as further described herein, and is in particular such that said VL domain is capable of binding to the peptide backbone of a glycosylated protein that is present or expressed on the surface of a cancer cell (and more in particular capable of binding to the peptide backbone of a glycosylated protein that is present or expressed on the surface of a cancer cell, where said glycosylated protein contains or presents a STn antigen) and/or such that said VL domain confers upon the antibody the ability to specifically bind to a protein that is present or expressed on the surface of a cancer cell (where, in particular, said protein is glycosylated and more in particular glycosylated such that it contains or presents a STn antigen).
The antibodies according to the preceding aspects can be as further described herein. Also, further aspects of the invention relate to nucleotide sequences that encode an antibody according to one of the preceding aspects, to host cells and/or host organisms that express or can be used to produce an antibody according to one of the preceding aspects, to pharmaceutical compositions that contain at least one antibody according to one of the preceding aspects, and to uses of an antibody according to one of the preceding aspects; all of which are preferably as further described herein.
DESCRIPTION OF THE INVENTION
Definitions and abbreviations
Unless specifically stated, as used herein, the term "nucleic acid" encompasses double as well as single-stranded nucleotide molecules. Nucleic acid sequences, when provided, are listed in the 5' to 3' direction, unless stated otherwise.
The term "homology", "similarity" or "sequence identity" between two proteins is determined by comparing the amino acid sequence and its conserved amino acid substitutes of one protein sequence to the second protein sequence. Similarity may be determined by procedures which are well-known in the art, for example, a BLAST program (Basic Local Alignment Search Tool at the National Center for Biological Information). Likewise, homology, similarity or sequence identity between two nucleic acid sequences may be determined by procedures which are well-known in the art, for example, a BLAST program (Basic Local Alignment Search Tool at the National Center for Biological Information).
"Amino acid residue substitution" at a specific position means substitution with any amino acid different from the native amino acid residue that is present at that specific position. A conservative amino acid substitution replaces an amino acid with another amino acid that is similar in size and chemical properties such that the substitution has no or only minor effect on protein structure and function; meanshile a nonconservative amino acid substitution replaces an amino acid with another amino acid that is dissimilar and thereby is likely to affect structure and function of the protein.
As used herein, the term "antibody" will be understood to include proteins having the characteristic two-armed, Y-shape of a typical antibody molecule as well as one or more fragments of an antibody that retain the ability to specifically bind to an antigen. Exemplary antibodies include, but are not limited to, a monoclonal antibody, a polyclonal antibody, a bi-specific antibody, a multispecific antibody, a grafted antibody, a human antibody, a humanized antibody, a synthetic antibody, a chimeric antibody, a camelized antibody, a single-chain Fvs (scFv) (including fragments in which the VL and VH are joined using recombinant methods by a synthetic or natural linker that enables them to be made as a single protein chain in which the VL and VH regions pair to form monovalent molecules, including single chain Fab and scFab), a single chain antibody, a Fab fragment (including monovalent fragments comprising the VL, VH, CL, and CHI domains), aF(ab')2 fragment (including bivalent fragments comprising two Fab fragments linked by a disulfide bridge at the hinge region), a Fd fragment (including fragments comprising the VH and CHI fragment), a Fv fragment (including fragments comprising the VL and VH domains of a single arm of an antibody), a single-domain antibody (dAb or sdAb) (including fragments comprising a VH domain), an isolated complementarity determining region (CDR), a diabody (including fragments comprising bivalent dimers such as two VL and VH domains bound to each other and recognizing two different antigens), a fragment comprised of only a single monomeric variable domain, disulfide-linked Fvs (sdFv), an intrabody, an anti-idiotypic (anti-Id) antibody, or ab antigen-binding fragments thereof.
As recognized by a person skilled in the art, the term VL refers to antibody variable domain, light chain, and the term VH refers to antibody variable domain, heavy chain.
The term "Tn antigen" as used herein refers to the monosaccharide structure N- acetylgalactosamine (GalNAc) linked to serine (Ser) or threonine (Thr) on a peptide backbone by a glycosidic bond (i.e. GalNAcal-O-Ser/Thr). The initials stand for Thomsen-nouveau. Tn antigen is expressed in most carcinomas. The term "Tn- carbohydrate epitope" (or "Tn epitope") as used herein refers to the GalNac part of the Tn antigen. The term "mono-Sn" refers to one Tn moiety (i.e. one GalNAc). The term "bis-Tn" refers to two Tn moieties (i.e. two GalNAc).
The term "STn antigen" as used herein refers to a sialyl-Tn antigen, formed by elongation of the Tn antigen with sialic acid (Neu5Ac(a2-6)GalNAc), still linked to serine (Ser) or threonine (Thr) (i.e. Neu5Aca2-6GalNAcal-O-Ser/Thr). Such STn antigen is also common on cancer tumor cells. Both Tn and STn may have additional modifications, such as phosphorylation, acetylation, methylation, and sulfonation. The term "STn- carbohydrate epitope" (or "STn epitope") as used herein refers to the Neu5Ac(a2- 6)GalNAc part of the Tn antigen. The term "mono-STn" refers to one STn moiety (i.e. one Neu5Ac(a2-6)GalNAc). The term "bis-STn" refers to two STn moieties (i.e. two Neu5Ac(a2-6)GalNAc).
The term glycoprotein generally refers to proteins which contain oligosaccharide chains covalently attached to amino acid side-chains. The term "glycoprotein" as used herein refers to a protein or a peptide thereof which contains a carbohydrate moiety (preferably a Tn or STn epitope) covalently attached to an amino acid residue (such as serine, threonine, or tyrosine; or any non-natural amino acid derivatives thereof such as replacing O with S) of said protein or peptide thereof. The term "combotope" as used herein refers to the combination of a carbohydrate epitope and a peptide epitope being recognized by an antibody. The two epitopes form a common epitope, the "combotope", which is different from each of the two epitopes from which it is composed. Of particular relevance for the present invention are combotopes where the peptide epitope is associated with a Tn epitope or a STn epitope; the peptide epitope may be associated with one or two Tn or STn epitopes. The combotope may be contiguous or discontiguous - i.e. the carbohydrate epitope is directly attached to the petide epitope by covalent bond (contiguous), or the carbohydrate epitope is attached to an amino acid residue located 1, 2, 3, 4, 5, 6, 7, 8, 9, or 10 amino acid residues up or down stream of the peptide epitope (discontiguous). Preferably, the carbohydrate epitope is covalently attached to the peptipe epitope.
The carbohydrate epitope may also be a combination of two Tn or STn epitopes (termed bis-Tn and bis-STn) or one Tn and one STn epitope, on adjacent amino acids in said peptide sequence in said glycoprotein. Adjacent means separated by 0 to 2 amino acids.
Antibodies which are refered to as "combotope binder" recognize and bind the combination of both the carbohydrate epitope and a peptide epitope of the glycoprotein. Such antibodies are also referred to as "combotope antibodies".
In the present application, the term "hapten" refers to the carbohydrate moiety(ies) (Tn or STn) independent of the peptide/protein. Antibodies which are referred to as "hapten binders" will bind Tn or STn epitopes independent of the peptide/protein carries - hence, they are therefore not specific for the combination of both the carbohydrate epitope and a peptide epitope of the glycoprotein (i.e. not a combotope binder).
The phase "domain which binds..." as used herein should be understood as "domain which is suitable for binding...", "domain which is capable of binding...", and/or "domain which is prepared for binding...".
Description of figures
Figure 1: Schematic illustration of an embodiment of the invention: Tn-template antibody library for identifying Tn-combotope antiobodies. Phage display of a library of scFv antibodes. Each scFv in the library has a Tn-binding VH domain. The library is screened for Tn-peptide specific scFv by biopanning using Tn-peptides.
Figure 2: Illustration of the prepararion of a phage display library of the present invention. 1) mRNA isolation from mouse spleen. 2) cDNA synthesis with reverse transcriptase using random hexamers. 3) PCR amplification from cDNA template to obtain the VH domain using a specific set primers, and the repertoire of VL domains using a mix of VL primesr. 4) PCR assembly of VL-domain repertoire and specific VH- domain using 5' phosphorylated outer primers. 5) Rolling circle amplification where phosphorylated scFv genes are ligated into circular DNA, dsDNA is denatured, random hexamers are annealed and Phi29 polymerase amplifies the circular fragments into long linear concatemers. 6) Amplified extended scFv genes are digested by sfil restriction enzyme and ligated to Sfil and rSAP treated phagemid vector pAKlOO. 7) Pool of phagemids containing scFv genes are electroporated to TGI E. coli cells. 8) Growth of Bacterial Library containing the different phagemids and infection with helper phage VCSM13 to produce complete phages displaying scFvs on their pill coat protein
Figure 3: X-ray of G2D11 scFv with APGS*T*AP peptide (where * denotes a GalNac residue) showed the interaction points of VH with the glycan structure. Key interaction points with two adjacent GalNac residues unclude His32H, Ala33H, His35H in CDR1; His40H, Ser52H, Asn55H, Asp57H in CDR2; and Ser99H in CRD3.
Figure 4: VH domain sequence alignment of G2D11 with other known VH-domains. Conserved amino acid residues are indicated by arrows (H32, A33, H35, Y50, S52, N55, D57 AND S99)
Figure 5: Phage and sequence enrichment after each round (Round 1, 2, and 3) of biopanning for bisTn-MUCl. Polyclonal phage ELISA confirmed phage enrichments for bis Tn-MUCl target peptide. bisTn MUC1 = peptide no 1 in Table 1; Tn MUC1 = peptide no 3 in Table 1; SA= negative control.
Figure 6. Confirmation of MUC1 scFv and selected clones for expression by dot blot analysis of scFv overnight expression in the 96 well format by His-tag detection.
Figure 7. MUC1 monoclonal ELISA. Screening of monoclonal scFvs on MUC1 target and control peptides.
Figure 8. MUC1 scFvs binding assays and kinetic affinities. (A) scFv titration at fixed concentration of MUC1 target peptide (peptide no 1 in Table 1). (B) scFv titration at fixed concentration of IgA hinge region control peptide (peptide no 8). Each data point is the mean value of three independent experiments. (C) Representative histograms of cell binding at 1.25 pg/ml of A3, D2 and D3 scFvs along with 5E5 mAb on MDA-MB-231 WT and COSMC KO cells. Flow cytometry experiments were repeated three times. Figure 9. MUC1 scFv titration on MUC1 peptides 2, 3, 4 and 5. MUC1 scFvs titrated on different MUC1 glycopeptides and unglycosylated MUC1. Asterisk denotes a glycosylation site.
Figure 10. MUC1 scFvs biological evaluation with flow cytometry. (A) Negative binding of MUC1 scFvs on HEK293 cells as a negative control cell line. (B) MCF7 cells at 1.25 pg/mL. C) Representative example of concentration dependent binding on MDA-MB-231 WT and COSMC KO cells. scFv D3 is shown at 4-fold dilution starting from 5 pg/mL.
Figure 11. X-ray of 5E5 scFv with APGST*AP peptide (asterisk denotes a GalNac residue). Tyr98L and the conserved TyrlOOL are the key amino acids in recognition of the peptide backbone.
Figure 12. Heat map of binding of scFv D3, scFv A4, scFv 5E5, scFv 2D9Chi, and scFv G2D11 to Tn-glycopeptides from Table 5. The glycopeptides were printed on a microarray chip. For simplicity, the heat map shows amino acids 9-19 of the peptides in Table 5. The relative fluorescence units (RFU) as shown as heat map. Tn-glycosylation sites are bold and underlined. Substitutions with Ala are marked as bold.
Figure 13. Polyclonal phage enrichment between the three rounds of selection against bisTn-CD43 target peptide (peptide no. 9 in Table 1) and control peptides (peptide no. 10 in Table 1).
Figure 14. CD43 monoclonal ELISA. Screening of monoclonal scFvs on CD43 target and control peptides.
Figure 15. CD43 scFvs binding assays. (A) Eight scFvs were were titrated on bisTn- CD43 target peptide (peptide no. 9 in Table 1). (B) ScFv titration on IgAl hinge region control glycopeptide (peptide no. 8 in Table 1) showed A7, D3 cross reactivity to IgA while Al and F4 showed weaker binding to IgAl. Each dot represents the mean value of three independent experiments. (C) Representative histograms of Al, D7, Hl and H2 scFvs at 1.25 pg/mL tested on Jurkat cells before and after neuraminidase treatment. Flow cytometry experiments were repeated three times.
Figure 16. CD43 scFvs biological evaluation with flow cytometry. Concentration dependent binding of Al scFv as a representative example on HEK293 cells and Jurkat cells, before and after neuraminidase treatment. scFv was 4-fold diluted starting from 5 pg/mL. Figure 17. CD43 x-ray structure with GAS*T*GSP peptide reveales the importance of Tyr99L as key interaction point with the peptide backbone.
Figure 18. Alignment of bisTn binder G2D11 VH with monoTn binder 3F1 VH to identify amino acid residues relevant for shifting to anti bisSTn.
Figure 19. Microaray data for binding of G2D11, 3F1, and mutants (M l-4) comprising selected mutations of VH-G2D11 to glycopeptides 1 (bisTnMUCl), 11 (bisSTnMUCl), 12 (monoSTnMUCl) and 4 (unglycosylated control).
Figure 20. Microaray data for binding of STnMUCl-D4, D3, C7 scFv to glycopeptides 1 (bisTnMUCl), 11 (bisSTnMUCl), and 12 (monoSTnMUCl).
Figure 21. ELISA titration screening of the humanized scFvs on coated Tn-MUCl and other Tn-proteins. (A) D3LlHlscFv, (B) D3LlH2scFv, (C) D3L2H3scFv, (D) D3L3H4scFv, (E) D3L4H5scFv, and (F) parental mouse D3. MUCl = mucin 1, MUC21 = mucin 21, GPNMB= Transmembrane Glycoprotein NMB, EGFR= Epidermal growth factor receptor, VVL= Vicia Villosa Lectin used for detecting Tn on the Tn-proteins.
Figure 22. Illustration of (A) bisTnMUCl, (B) monoTn(Thr)MUCl, and (C) monoTn(Ser)MUCl.
Figure 23. Elisa titration of different target Tn-peptides detected with mouse D3 (lug/mL) scFvs on streptavidin coated plates. Column 1 : Biotin-2OEG-2OEG- HSSSTIPTPA(MUC13), column 2: Biotin-2OEG-2OEG-HSSSTIPIPT(MUC13), column 3: Biotin-2OEG-2OEG-SESITNVNSL(MUC13), column 4: Biotin-2OEG-2OEG-
ITASSPNDGL(MUC13), column 5: Biotin-2OEG-2OEG-MSPHEDNQz(MYC13), column 6: Biotin-2OEG-2OEG-DNQSSGPPTG(MUC13), column 7: Biotin-2OEG-2OEG-
LHNTSFCLCL(MUC13), column 8: Biotin-2OEG-2OEG-YNSSTCKKGK(MUC13), column 9: Biotin-2OEG-2OEG-IRSSSSNFLN(MUC13), column 10: Biotin-2OEG-2OEG-
CVASSLKCPD(MUC13), column 11 : Biotin-2OEG-2OEG-SITSTGLTSP(MUC4), column 12: Biotin-2OEG-2OEG-APGSTAPPAH(MUC1).
Detailed description of the invention
The present invention concerns a new antibody concept technology for simple and rapid development of antibodies targeting Tn- and STn- glycosylation sites of any glycoprotein site of choice. Specifically, the invention concerns antibody libraries which can be screened for antibodies which have improved specificity due to their specificity towards a combination of an epitopes on a carbohydrate part of a glycoprotein and an epitope on a peptide backbone in said glycoprotein which is associated with the carbohydrate epitope. The combined epitope is termed a "combotope". A non-limiting embodiment of the invention is schematically illustrated in Figure 1.
The present invention is especially useful in the generation of therapeutics for cancer treatment and diagnostics.
I. Combotope antibodies
The present invention provides combotope antibodies which have high specificity and high binding efficiency to their target glycopeptide due to their combined specificity towards both the carbohydrate epitope as well as the peptide backbone epitope associated with the carbohydrate epitope of the glycoprotein target. In one preferred embodiment, the present invention provides a combotope antibody for targeting tumor cells carrying said glycoprotein on their surface.
In one aspect, the invention provides an antibody for targeting tumor cells, said antibody comprising two antibody domains, where the first antibody domain binds a carbohydrate epitope of a glycoprotein of the tumor cell and the second antibody domain binds a peptide epitope of said glycoprotein of said tumor cell, where said antibody specifically binds both epitopes (as a common epitope) as compared to only binding one of said epitopes.
In one embodiment, the first antibody domain is a VH domain and the second antibody domain is a VL domain. In another embodiment, both the first and second antibody domains are VH domains, but different from one another.
In one preferred embodiment, the antibodies disclosed herein are scFv antibodies, comprising a VH domain and a VL domain, where both domains are present in a single polypeptide chain. In some embodiments, the Fv polypeptide further comprises a polypeptide linker between the VH and VL domains allowing the scFv to form the desired structure for antigen binding. In one embodiment, the linker is selected from (GGGGS)n where in is 1, 2, 3, 4, 5, or 6; e.g. (GGGGS)4 or (GGGGS)s. Other linkers could be used as an alternative to the exemplified (GGGGS)n-linker. The skilled person would know how to select such linkers.
In one aspect, the invention provides an antibody for targeting tumor cells, said antibody comprising (i) a VH domain binding a carbohydrate epitope of a glycoprotein of the tumor cell and (ii) a VL domain binding a peptide epitope of said glycoprotein of said tumor cell, where said antibody is selected to specifically binding both epitopes (as a common epitope) as compared to only binding one of said epitopes.
In one embodiment, the invention provides an antibody for targeting a glycoprotein, such as a specific glycoprotein on a tumor cell. The glycoprotein comprises a carbohydrate epitope and a peptide epitope. The carbohydrate epitope is a short truncated O-glycan, such as Tn or STn. The peptide epitope is associated with the carbohydrate epitope, such as directly attached by virtue of being chemically linked or being in close proximity of one another by virtue of the structural configuration of the glycoprotein. The antibody of the present invention is characterized by its ability to specifically bind both epitopes (as a common epitope), as compared to only binding one of said epitopes. In one embodiment, the antibody comprises (i) a VH domain characterized by (a) being suitable for binding a carbohydrate epitope of a glycoprotein of a tumor cell and (b) not binding the peptide epitope of said glycoprotein of said tumor cell and (ii) a VL domain characterized by being suitable for binding the peptide epitope of said glycoprotein of said tumor cell. In one embodiment, the VL-domain is further characterized by not binding the carbohydrate epitope of said glycoprotein. In a specific embodiment, the antibody only binds a combination of carbohydrate and peptide epitopes, i.e. both epitops need to be present for binding of the antibody.
In one embodiment, the present invention provides an antibody which binds a tumor cell, wherein the antibody comprises
(i) a VH-domain which binds a carbohydrate epitope of a glycoprotein of a tumor cell, and
(ii) a VL-domain which binds a peptide epitope of said glycoprotein of said tumor cell, wherein the peptide epitope of the glycoprotein of the tumor cell is associated with the carbohydrate epitope of said glycoprotein of said tumor cells, and wherein the antibody is specific for the combination of said carbohydrate epitope and said peptide epitope on said glycoprotein of said cancer cell.
In one embodiment, the present invention provides an antibody which binds a tumor cell, wherein the antibody comprises (ii) a VL-domain which binds a peptide epitope of a glycoprotein of said tumor cell and (i) a VH-domain which binds a carbohydrate epitope of said glycoprotein of said tumor cell and which does not contribute to or interfere with binding said peptide epitope of said glycoprotein of said tumor cell, i.e. the VH-domain binding is not affected/influenced by the presense of any peptide epitope; and wherein the peptide epitope of the glycoprotein of the tumor cell is associated with the carbohydrate epitope of said glycoprotein of said tumor cells, and wherein the antibody is specific for the combination of said carbohydrate epitope and said peptide epitope on said glycoprotein of said cancer cell.
In one embodiment, the present invention provides an antibody which binds a tumor cell, wherein the antibody comprises (ii) a VL-domain which binds a peptide epitope of a glycoprotein of said tumor cell and (i) a VH-domain which exclusively binds a carbohydrate epitope of said glycoprotein of said tumor cell, i.e. which does not contribute to or interfere with binding peptide epitope of said glycoprotein of said tumor cell; and wherein the peptide epitope of the glycoprotein of the tumor cell is associated with the carbohydrate epitope of said glycoprotein of said tumor cells, and wherein the antibody is specific for the combination of said carbohydrate epitope and said peptide epitope on said glycoprotein of said cancer cell.
As disclosed here, the antibody of the present invention is specific for the combination of the carbohydrate epitope, which is a short truncated O-glycan, and the peptide epitope of a glycopeptide of a tumor cell, such as on the surface of the tumor cell. The term "specific" in this regard refers to the specific antibody being highly selective for a particular glycoprotein, exhibiting strong binding and recognition for that glycoprotein, while displaying minimal or no binding to other types of glycoproteins. Specificity may be expressed by determining the binding affinity of the antibody to the glycoprotein, by using biophysical techniques as recognized by a person skilled in the art. The antibodies of the present invention recognize both the carbohydrate moiety and the peptide sequence of the glycoprotein making them very specific.
In one embodiment, the peptide epitope recognized by the VL-domain is part of a specific glycoprotein on the surface of a specific type of cancerous cells.
In addition, the combotope antibodies of the invention is selected to be specific for the combined glycoprotein epitope, i.e. the carbohydrate epitope in combination with the peptide epitope (the combotope), and not for each of the epitopes individually.
In one embodiment, the present invention provides an antibody which binds one or more tumor cells, wherein the one or more tumor cells comprises a carbohydrate epitope, preferably a short truncated O-glycan carbohydrate epitope, associated with a peptide epitope, wherein the antibody is specific against both the carbohydrate epitope and the peptide epitope. In one embodiment, the specific antibody does not bind specifically to combinations other than the specific combotope, and does not bind (or binds less effectively) the carbohydrate epitope or peptide epitope separately.
In one embodiment, the present invention provides an antibody which binds a tumor cell, wherein the tumor cell comprises a carbohydrate epitope, preferably a short truncated O-glycan, associated with a peptide epitope, wherein the antibody is specific for the combination of said carbohydrate epitope and said peptide epitope by virtue of the antibody comprising
(i) a VH-domain which binds the carbohydrate epitope, and
(ii) a VL-domain which binds the peptide epitope.
Preferably the carbohydrate epitope is part of a glycoprotein and the peptide epitopes is part of said same glycoprotein. The carbohydrate and peptide epitopes of the tumor are preferably surface displayed in order to facilitate recognition by the antibody.
In one preferred embodiment, the carbohydrate epitope to which the VH-domain binds is selected from mono-Tn, bis-Tn, mono-STn, bis-STn, and/or a combination of monoTn and monoSTn. The carbohydrate epitope may comprise or consist of one or two adjacent short truncated O-glycans, i.e. Tn, TrnTn, STn, STrnSTn. Hence, in one preferred embodiment, the present invention provides an antibody which binds a tumor cell, wherein the antibody comprises
(i) a VH-domain which binds a monoTn, bisTn, monoSTn, bisSTn, and/or monoTn+monoSTn carbohydrate epitope of a glycoprotein of the tumor cells, and
(ii) a VL-domain which binds a peptide epitope of said plycoprotein of said tumor cells, wherein the peptide epitope of the glycoprotein of the tumor cell is associated with the carbohydrate epitope of said glycoprotein of the tumor cell, and wherein the antibody is specific for the combination of said carbohydrate epitope and said peptide epitope of said glycoprotein of said cancer cell.
As disclosed herein, the target peptide epitope of the cell is associated with a carbohydrate epitope. The term "associated with" preferably refers to the peptide epitope and carbohydrate epitope being a continuous epitope, i.e. where the carbohydrate is directly attached to the peptide by virtue of being chemically linked, such as covalently linked. In another emodiment, the peptide epitope and carbohydrate epitope may be a discontinuous epitope, where the "associated with" still refers to the peptide epitope and carbohydrate epitope being in close proximity of one another, but by virtue of the structural configuration of the molecule facilitating this. Hence, a discontinued epitope is where the amino acid hosting the attached carbohydrate epitope is not part of the peptide epitope.
In one embodiment, the present invention provides a tumor cell binding antibody, comprising
(i) a VH-domain which binds a carbohydrate epitope of a glycoprotein of the tumor cell, preferably the carbohydrate epitope is one or two short truncated O- glycan(s), and
(ii) a VL-domain which binds a peptide epitope of said glycoprotein of said tumor cell, wherein the antibody is specific for the combination of said carbohydrate epitope and said peptide epitope of said glycoprotein, but not specific for other combinations or one of the epitopes alone.
In one embodiment, the present invention provides a tumor cell binding antibody, comprising
(i) a VH-domain which binds a Tn- and/or STn-carbohydrate epitope of a glycoprotein of the tumor cell, preferably the carbohydrate epitope is one or two short truncated O-glycan(s), and
(ii) a VL-domain which binds a peptide epitope of said glycoprotein of said tumor cell, wherein the antibody is specific for the combination of said carbohydrate epitope and said peptide epitope of said glycoprotein, but not specific for other combinations or one of the epitopes alone.
The carbohydrate epitope is composed of one or two short truncated O-glycan(s), where the a short truncated O-glycans are selected from Tn and STn.
A Tn or a STn epitope may be formed by one or two Tn or STn moieties. In one embodiment, the Tn or STn epitope which interacts with the VH domain is only one Tn or one STn moiety, respectively. In one preferred embodiment, the Tn or STn epitope which interacts with the VH domain is only one Tn or one STn moiety, and said Tn or STn epitope is covalently attached to an amino acid residue, wherein said amino acid residue is part of the peptide epitope which interacts with the VL-domain.
In another embodiment, the Tn or STn epitope is formed by two Tn or STn moieties, respectively, which both interacts with the VH domain. In one preferred embodiment, the Tn or STn epitope which interacts with the VH domain is two Tn or one STn moieties, and said two Tn or STn moieties are covalently attached to two separate amino acid residue, wherein said amino acid residues are part of the peptide epitope which interacts with the VL-domain. In one such embodiment said two separate amino acid residues are adjacent to each other; in another embodiment said two separate amino acid residies are spaced apart by 1, 2, 3 or 4 other amino acid residues.
As disclosed herein, the VL-domain of the combotope antibody binds a peptide epitope of a glycoprotein of a tumor cell. In one embodiment, the peptide epitope is 2, 3, 4, 5, 6, 7, 8, 9, 10, 11, or 12 amino acid residues, preferably 2, 3, or 4 amino acid residues, most preferably 4 amino acid residues. In one embodiment, the peptide epitope which interacts with the VL domain is between 2-4, between 4-6, between 6-8, between 8-10, between 10-12 amino acid residues, such as between 2-12, between 2-10, between 2- 8, between preferably 2-6 amino acid residues, most preferably 2-4 amino acid residues.
In a most preferred embodiment, the antibodies of the present invention are different from the prior art antibodies 5E5, 5F7, and 2D9.
5F7 is disclosed in US11161911B2. 5E5 is disclosed in W02008/040362 and US2021060070A1, and further in the scientific literature Macias-Leon et al 2020; Tarp et al 2007; and Blixt et al 2010. 2D9 is disclosed in the scientific litterature Sorensen et al 2006; Tarp et al 2007; and Blixt et al 2010.
In one embodiment, the antibodies of the present invention are different from the prior art antibodies 5E5, 5F7, and 2D9 - hence, the antibodies of the present invention do not comprise VL and VH domain combinations as disclosed here:
In one embodiment, the amino acid sequence of the antibodies of the present invention do not comprise an amino acid sequence combination selected from SEQ ID NO. 3 + 4, SEQ ID NO. 5 + 6, and SEQ ID NO. 7 + 8.
Preferably, the antibody of the present invention is a humanized antibody, such as prepared by the method of Clavero-Alvarez et al 2018. "Humanized" forms of nonhuman antibodies can be chimeric antibodies that contain minimal sequence derived from the non-human antibody. A humanized antibody is generally a human antibody (recipient antibody) in which selected residues in the non-human antibody (donor antibody) have been replaced. The donor antibody can be any suitable non-human antibody, such as a mouse, rat, rabbit, chicken, or non-human primate antibody having a desired specificity, affinity, or biological effect. For the present invention, the donor antibody is preferably identified by screening an antibody library of the present invention. In some instances, selected framework region residues of the recipient antibody are replaced by the corresponding framework region residues from the donor antibody. Humanized antibodies may also comprise residues that are not found in either the recipient antibody or the donor antibody. In some instances, these modifications are made to further refine antibody performance. The skilled person would be familiar of methods to transform combotope antibodies of the present invention into humanized combotope antibodies.
In a further ascpect, the present invention provides nucleic acid sequences encoding the antibodies according to the present invention as disclosed herein.
I.i Tn-combotopes
As disclosed herein, the VH domain of the combotope antibody binds the carbohydrate epitope, preferably a mono-Tn, bis-Tn, mono-STn or bis-STN carbohydrate epitope.
In one preferred aspect, the VH domain of the combotope antibody binds a Tn- carbohydrate epitope, such a mono-Tn or bis-Tn.
The inventors of the present invention surprisingly made the following discovery: The VH-domain of antibody G2D11 (SEQ ID NO. 1) was by structural characterization in the presence of the bis-Tn-MUCl peptide APGS*T*AP (where * denotes a GalNAc moiety) found to recognize the two GalNAc moieties, but did not recognize the peptide sequence on which they were attached (see examples 1).
SEQ ID NO. 1 (VH-domain G2D11): QVQMQQSDAELVKPGASVKISCKASGYIFADHAIHWVKRKPEQGLEWIGYISPGNDDIKYNEKF KGKATLTADKSSSTAYMQLNSLTSEDSAVYFCKRSLPGTFDYWGQGTTLTVSS
The novel combotope antibodies disclosed herein are partially based on these structural observations - i.e. that the VH-chain of G2D11 provides recognition support for the glycoside part of the antigen, without being affected/influenced by the peptide part of the glycoprotein and did not bind to this peptide. The further development, as disclosed herein, specifies a combotope antibody, where the VL-domain provides specific binding of the peptide epitope of the combotope and the VH-chain provides recognition support for the carbohydrate part of the glycoprotein. As disclosed in the background section, several anti-Tn antibodies are known in the art, and many of these share the same germline sequences as the G2D11 (at least the essential sequences). But until now it was not know that the VH domain of G2D11 provides exclusive support for the recognition of the Tn carbohydrate epitope, without binding to the peptide epitope. Hence, the prior art anti-Tn antibodies may comprise same germline sequences as the G2D11 responsible for the Tn-binding VH domain, but with no or an unknown binding contribution to the peptide/protein carrier - hence, they are unspecific Tn-binding antibodies, and therefore not combotope antibodies according to the present invention. On the contrary, combotopes of the present invention are screened for by use of the antibody library of the present invention, as further disclosed herein, to obtain antibodies specific for a specific glycoproteins of interest.
In one embodiment, the VH-domain of the combotope antibody of the present invention is a G2D11-Iike VH-domain. In another embodiment, the amino acids of the VH-domain of the combotope antibody of the present invention resemble the G2D11-Iike VH-domain in their structural conformation.
G2D11 tolerates binding of any combination Tn-Thr/Tn-Ser, Tn-Ser/Tn-Thr, Tn-Ser/Tn- Ser, Tn-Thr/Tn-Thr. As demonstrated in Example 1 herein, with reference to SEQ ID NO. 1 (VH domain of G2D11), amino acid residues H32, A33, H35, Y50, and S99 are key residues for the binding of the VH domain to one of the GalNAc moiety of bis-Tn-MUCl on the peptide, while amino acid residues S52, N55, and D57 are key residues for the binding of the VH domain to the other GalNAc moiety of bis-Tn-MUCl on the peptide.
In one embodiment, the VH-domain of the combotope antibody of the present invention comprises an amino acid sequence having at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 1, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 1; and comprises the amino acid residues H32, A33, H35, Y50, and S99, with respect to SEQ ID NO. 1. Said combotope antibody will provide recognition support for mono-Tn epitopes.
Hence, in one embodiment, combotope antibodes for targeting tumor cells are provided, said antibody comprises a VH and a VL domain; wherein the VH domain of the antibodyis a Tn-binding domain and the amino acid sequence of the VH-domain has at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 1, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 1, wherein the amino acid sequence comprises the amino acid residues H32, A33, H35, Y50, and S99, with respect to SEQ ID NO. 1; and wherein the VL domain of the antibody binds a peptide backbone epitope associated with the Tn-carbohydrate on the tumor cell. The sequence of the VH domain is preferably such that there is no binding to the peptide epitope.
In one embodiment, the VH-domain of the combotope antibody of the present invention comprises an amino acid sequence having at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 1, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 1; and comprises the amino acid residues S52, N55, and D57, with respect to SEQ ID NO. 1. Said combotope antibody will provide recognition support for mono-Tn epitopes.
Hence, in one embodiment, combotope antibodes for targeting tumor cells are provided, said antibody comprising a VH and a VL domain; wherein the VH domain of the antibody is a Tn-binding domain and the amino acid sequence of the VH-domain has at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 1, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 1, wherein the amino acid sequence comprises the amino acid residues S52, N55, and D57, with respect to SEQ ID NO. 1; and wherein the VL domain of the antibody binds a peptide backbone epitope associated with the Tn- carbohydrate on the tumor cell. The sequence of the VH domain is preferably such that there is no binding to the peptide epitope.
As demonsted herein, sequence alignment of amino acid seqences of VH-domains of selected Tn-binding mAbs, including the VH-domain of G2D11, showed conserved amino acids in the CDR1, CDR2 and CDR3 regions, related to binding the GalNac (see example 1.2). Specifically, with reference to SEQ ID NO. 1, amino acid residues H32, A33, and H35 in CDR1 should preferably be conserved for the VH domain of the present invention; further, with reference to SEQ ID NO. 1, amino acid residues Y50, S52, N55, and D57 in the CDR2 should preferably be conserved for the VH domain of the present invention; and further, with reference to SEQ ID NO. 1, amino acid residue S99 in the CDR3 should preferably be conserved for the VH domain of the present invention.
In one embodiment, the VH-domain of the combotope antibody comprises an amino acid sequence having at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 1, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 1; and comprises the amino acid residues H32, A33, H35, Y50, S52, N55, D57, and S99, with respect to SEQ ID NO. 1. Said combotope antibody will also provide recognition support for the bis-Tn epitope. In one embodiment, the amino acid sequence of the VH-domain of the combotope antibody has at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 1, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 1; and comprises the amino acid residues H32, A33, H35, Y50, S52, N55, D57, and S99, with respect to SEQ ID NO. 1.
In one embodiment, the VH domain of the combotope antibody comprises an amino acid sequence having at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 1, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 1; and in pairwise alignment with SEQ ID NO. 1, the amino acid sequence of the VH-domain comprises amino acid residues histidine (H), alanine (A), histidine (H), tyrosine (Y), serine (S), asparagine (N), aspartic acid (D), and serine (S) at positions corresponding to amino acid position H32, A33, H35, Y50, S52, N55, D57, and S99 of SEQ ID NO. 1, respectively. The pairwise sequence alignment is performed using scoring matrix: blosum62, gap opening penalty: 10, and gap extension penalty 0.2.
In one embodiment, in pairwise alignment with SEQ ID NO.: 1, the amino acid sequence of the VH-domain of the combotope antibody comprises amino acid residues histidine (H), alanine (A), histidine (H), tyrosine (Y), serine (S), asparagine (N), aspartic acid (D), and serine (S), at positions corresponding to amino acid position H32, A33, H35, Y50, S52, N55, D57, and S99, of SEQ ID NO. 1, respectively; and the amino acid sequence of the VH domain has at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 1, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 1.
Hence, in one preferred embodiment, antibodes for targeting tumor cells are provided, said antibody comprising a VH and a VL domain; wherein the VH domain of the antibody is a Tn-binding domain and the amino acid sequence of the VH-domain has at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 1, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 1, wherein the amino acid sequence comprises the amino acid residues H32, A33, H35, Y50, S52, N55, D57, and S99, with respect to SEQ ID NO. 1; and wherein the VL domain of the antibody binds a peptide backbone epitope associated with the Tn-carbohydrate on the tumor cell. The sequence of the VH domain is preferably such that there is no binding to the peptide epitope. Combotope antibodies of the present invention as described herein, such as the Tn- combotopes disclosed herein, comprise improved binding affinity to a specific antigen epitope termed a combotope comprised of two different epitopes on a glycoprotein. In some embodiments, the antibody comprises a binding affinity (e.g. kD) of between 100 nM to IpM, such as less than 100 nM, less than 10 nM, less than 1 nM, less than 100 pM, or even less than 10 pM.
In some embodiment, the combotope antibodies of the present invention are used to treat cancer. In some instances, the cancer is lung, head and neck squamous cell, colorectal, melanoma, liver, classical Hodgkin lymphoma, kidney, gastric, cervical, merkel cell, B-cell lymphoma, or bladder cancer. In a preferred embodiment, the cancer is a solid tumor.
In yet another aspect, provided herein are specific antibodies
In one embodiment, the antibody comprises
(i) a VH-domain having at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 1, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 1, and comprising the amino acid residues H32, A33, H35, Y50, S52, N55, D57, and S99, with respect to SEQ ID NO. 1, and
(ii) a VL domain having an amino acid sequence selected from of any one of SEQ ID NOs. 9-21.
In one embodiment the present invention provides an antibody comprising a VH domain having the amino acid sequence of SEQ ID NO. 1 and a VL domain having an amino acid sequence seleted from any one of SEQ ID NO. 9-21.
In one embodiment, the present invention provides antibodies, wherein the antibody comprises (i) a VH-domain having at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 1, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 1, and comprising the amino acid residues H32, A33, H35, Y50, S52, N55, D57, and S99, with respect to SEQ ID NO. 1, and (ii) a VL domain having an amino acid sequence selected from of any one of SEQ ID NOs: 9-21; and wherein the antibody is a monoclonal antibody, a polyclonal antibody, a bi-specific antibody, a multispecific antibody, a grafted antibody, a human antibody, a humanized antibody, a synthetic antibody, a chimeric antibody, a camelized antibody, a single-chain Fvs (scFv), a single chain antibody, a Fab fragment, a F(ab')2 fragment, a Fd fragment, a Fv fragment, a single-domain antibody, an isolated complementarity determining region (CDR), a diabody, a fragment comprised of only a single monomeric variable domain, disulfide-linked Fvs (sdFv), an intrabody, an anti-idiotypic (anti-Id ) antibody, or ab antigen-binding fragments thereof.
In one embodiment, an antibody comprising (i) a VH domain having an amino acid sequence of SEQ ID NO. 1 and (ii) a VL domain having an amino acid sequence of any one of SEQ ID NOs. 9-21 is used to treat cancer. In some instances, the cancer is lung, head and neck squamous cell, colorectal, melanoma, liver, classical Hodgkin lymphoma, kidney, gastric, cervical, merkel cell, B-cell lymphoma, or bladder cancer. In a preferred embodiment, the cancer is a solid tumor.
In another embodiment, an antibody comprising (i) a VH domain having an amino acid sequence of SEQ ID NO. 1 and (ii) a VL domain having an amino acid sequence of any one of SEQ ID NOs. 9-21 is used to diagnose a cancerous state. The antibody may be used for in vivo diagnosis or for identifying a cancerous state ex vivo in a tissue or cell sample taken from a patient.
Binding of the combotope antibody to a cancer target for diagnosis can be monitored or identified by way of known methods, such as for example labelling of the combotope antibodies and/or by use of labelled antibodies for binding of the combotope antibodies.
SEQ ID NO. 9 (VL-domain D3-TnMUCl) :
DYKDIQMTQSPSSLAVSVGEKVTMSCKSSQSLLYSSNQKNYLAWYQQKPGQSPKLLIYWASTRE SG VP D RFTG SGSGTD FTLTISS VKAE D LAVYYCQQYYSYP LTFG AGTKLE M KR
SEQ ID NO. 10 (VL-domain A3-TnMUCl) :
DYKDIVMTQSQKFMSTSVGDRVSITCKASQNVGTAVAWYQQKPGQSPKLLIYSASNRYTGVPDR FTGSGSGTDFTLTISNVQSEDLADYFCLQHWNYPLTFGGGTKLEIKR
SEQ ID NO. 11 (VL-domain D2-TnMUCl) :
DYKDIQMTQSHKFMSTSVGDRVSITCKASQDVGTAVAWYQQKPGQSPKLLIYWASTRHTGVPD RFTGSGSGTDFTLTISNVQSEDLADYFCLQHWNYPLTFGGGTKLEIKR
SEQ ID NO. 12 (VL-domain Ori-TnCD43) :
DYKDIQMTQSPASLSASVGETVTITCRASENIYSYLAWYQQKQGKSPQLLVYNAKTLAEGVPSRF SGSGSGTQFSLKINSLQPEDFGSYYCQHHYGTPYTFGGGTKLEIKR
SEQ ID NO. 13 (VL-domain H l-TnCD43) :
DYKDIVMTQSPSSLAVSVGEKVTMSCKSSQSLLYSSNQKNYLAWYQQKPGQSPKLLIYWASTRE SGVPDRFTGSGSGTDFTLTISSVKAEDLAVYYCQQYYSYPWTFGGGTKLEIKR
SEQ ID NO. 14 (VL-domain Al-TnCD43) :
DYKDIVMTQS PAS LSAS VG ETVTITCRAS E N IYSYLA WYQQ KQG KS PQ LLVYN AKTLAEG VPS RF SGSGSGTQFSLKINSLQSEDFGSYYCQHHYGTPYTFGGGTKLEIKR
SEQ ID NO. 15 (VL-domain F4-TnCD43) : DYKDIQMTQSPASLSASVGETVTITCRASENIYSYLAWYQQKQGKSPQLLVYNAKTLAEGVPSRF SGSGSGTQYSLKINSLQPEDFGSYYCQHFWSTPYTFGGGTKLEMKR
SEQ ID NO. 16 (VL-domain C5-TnCD43) :
DYKDVQMTQSHKFMSTSVGDRVSITCKASQDVSTAVAWYQQKPGQSPKLLIYWASTRHTGVPD RFTGSGSGTDYTLTISSVQAEDLALYYCQQHYSTPYTFGGGTKLEIKR
SEQ ID NO. 17 (VL-domain C5-TnCD43) :
DYKDIVMTQSHKFMSTSVGDRVSITCKASQDVGTAVAWYQQKPGQSPKLLIYWASTRHTGVPD RFTGSGSGTDFTLTISNVQSEDLADYFCQQYSSYPYTFGGGTKLEMKR
SEQ ID NO. 18 (VL-domain D3-TnCD43):
DYKDIVMTQS PSS LAVSAG E KVTMSCKSSQSLLNS RTRKNYLAWYQQKPGQS PKLLIYWASTRE SGVPDRFTGSGSGTDFTLTISNVQSEDLAEYFCQQYNSYPLTFGAGTKLEIKR
SEQ ID NO. 19 (VL-domain G3-TnCD43) :
DYKDVVMTQSQKFMSTSVRDRVSITCKASQNVGTAVAWYQQKPGQSPKLLIYSASYRYSGVPD
H FTGSG SGTD FTLTIS N VQS E D LAEYFCQQYYSYPYTFGG GTKLEI KR
SEQ ID NO. 20 (VL-domain D7-TnCD43) :
DYKDLVLTQSPSSLAVSVGEKVTMSCKSSQSLLYSSNQKNYLAWYQQKPGQSPKLLIYWASTRE SGVPDRFTGSGSGTDFTLTISSVKAEDLAVYYCQQYYSYPWTFGGGTKLEM KR
SEQ ID NO. 21 (VL-domain H2-TnCD43) :
DYKDIVMTQSPSSLAVSVGEKVTMSCKSSQSLLYSSNQKNYLAWYQQKPGQSPKLLIYWASTRE SGVPDRFTGSGSGTDFTLTISSVKAEDLAVYYCQQYYSYPYTFGGGTKLEIKR
Preferably, the selected antibodies of the present invention are a humanized antibody, such as mententioned above.
The following combinations of humanized VH and LV domains were found of particular interest: D3VL1 + D3VH 1, D3VL1 + D3VH2, D3VL2+D3VH3, D3VL3 + D3VH4, and D3VL4 + D3VH5.
> D3VL1 (SEQ ID NO. 159)
DIVMTQSPDSLAVSLGERATINCKSSQSLLYSSNQKNYLAWYQQKPGQAPKLLIYWASTRESGVP DRFSGSGSGTDFTLTISSLQAEDVAVYYCQQYYSYPLTFGAGTKLEMK > D3VH 1 (SEQ ID NO. 160) EVQLVQSGAEVKKPGASVKVSCKASGYIFADHAIHWVRQAPGQRLEWIGYISPGNDDIKYNQKF QGRVTLTADKSASTAYMELSSLRSEDSAVYFCKRSLPGTFDYWGQGTTLTVSS
> D3VL1 (SEQ ID NO. 159)
DIVMTQSPDSLAVSLGERATINCKSSQSLLYSSNQKNYLAWYQQKPGQAPKLLIYWASTRESGVP DRFSGSGSGTDFTLTISSLQAEDVAVYYCQQYYSYPLTFGAGTKLEMK > D3VH2 (SEQ ID NO. 161) EVQLVQSGAEVKKPGSSVKVSCKASGYIFADHAIHWVRRAPGQGLEWIGYISPGNDDIKYNEKF
KGRATLTADKSTSTAYMELSSLRSEDTAVYFCKRSLPGTFDYWGQGTTLTVSS
>D3VL2 (SEQ ID NO. 162)
DIVMTQSPDSLAVSLGEKATINCKSSQSLLYSSNQKNYLAWYQQKPGQPPKLLIYWASTRESGVP DRFSGSGSGTDFTLTISSLQAEDVAVYYCQQYYSYPLTFGGGTKVEIK >D3VH3 (SEQ ID NO. 163) EVQLVQSGAEVKKPGASVKVSCKASGYIFADHAIHWVRQAPGQRLEWIGYISPGNDDIKYSQKF QDKVTLTADKSASTAYMELSSLRSEDTAVYFCKRSLPGTFDYWGQGTTVTVSS
>D3VL3 (SEQ ID NO. 164)
DIQMTQSPSSVSASVGDRLTITCRSSQSLLYSSNQKNYLAWYQQKPGKAPKLLIYWASSLQSGV PSRFSGSGSGTDFTLTISSLKPEDFATYYCQQYYSYPLTFGQGTKVEIK >D3VH4 (SEQ ID NO. 165) EVQLVQSGAEVKKPGASVKVSCKASGYIFADHAIHWVRQAPGQRLEWIGYISPGNDDIKYSQEF QGRVTLTADKSASTAYMELSSLRSEDSAVYFCKRSLPGTFDYWGQGTTLTVSS
>D3VL4 (SEQ ID NO. 166)
DIQMTQSPDSLAVSLGERATINCKSSQSLLYSSNQKNYLAWYQQKPGQPPKLLIYWASTRESGVP DRFSGSGSGTDFTLTISSLQAEDVAVYYCQQYYSYPLTFGQGTKVEIK >D3VH5 (SEQ ID NO. 167) EVQLVQSGAEVKKPGASVKVSCKASGYIFADHAIHWVRQAPGQRLEWIGYISPGNDDIKYSQEF QGRVTLTADKSASTAYMELSSLRSEDSAVYFCKRSLPGTFDYWGQGTTLTVSS
In one embodiment the present invention provides a humanized antibody comprising a VH domain having the amino acid sequence of SEQ ID NO. 159 and a VL domain having an amino acid sequence of SEQ ID NO. 160. In one embodiment, a humanized antibody comprising a VH domain having the amino acid sequence of SEQ ID NO. 159 and a VL domain having an amino acid sequence of SEQ ID NO. 160 is used to treat cancer, e.g. a solid tumor cancer. In another embodiment, a humanized antibody comprising a VH domain having the amino acid sequence of SEQ ID NO. 159 and a VL domain having an amino acid sequence of SEQ ID NO. 160 is used to diagnose a cancerous state, e.g. by in vivo diagnosis or for identifying a cancerous state in a tissue or cell sample taken from a patient.
In one embodiment the present invention provides a humanized antibody comprising a VH domain having the amino acid sequence of SEQ ID NO. 159 and a VL domain having an amino acid sequence of SEQ ID NO. 161. In one embodiment, a humanized antibody comprising a VH domain having the amino acid sequence of SEQ ID NO. 159 and a VL domain having an amino acid sequence of SEQ ID NO. 161 is used to treat cancer, e.g. a solid tumor cancer. In another embodiment, a humanized antibody comprising a VH domain having the amino acid sequence of SEQ ID NO. 159 and a VL domain having an amino acid sequence of SEQ ID NO. 161 is used to diagnose a cancerous state, e.g. by in vivo diagnosis or for identifying a cancerous state in a tissue or cell sample taken from a patient.
In one embodiment the present invention provides a humanized antibody comprising a VH domain having the amino acid sequence of SEQ ID NO. 162 and a VL domain having an amino acid sequence of SEQ ID NO. 163. In one embodiment, a humanized antibody comprising a VH domain having the amino acid sequence of SEQ ID NO. 162 and a VL domain having an amino acid sequence of SEQ ID NO. 163 is used to treat cancer, e.g. a solid tumor cancer. In another embodiment, a humanized antibody comprising a VH domain having the amino acid sequence of SEQ ID NO. 162 and a VL domain having an amino acid sequence of SEQ ID NO. 163 is used to diagnose a cancerous state, e.g. by in vivo diagnosis or for identifying a cancerous state in a tissue or cell sample taken from a patient.
In one embodiment the present invention provides a humanized antibody comprising a VH domain having the amino acid sequence of SEQ ID NO. 164 and a VL domain having an amino acid sequence of SEQ ID NO. 165. In one embodiment, a humanized antibody comprising a VH domain having the amino acid sequence of SEQ ID NO. 164 and a VL domain having an amino acid sequence of SEQ ID NO. 165 is used to treat cancer, e.g. a solid tumor cancer. In another embodiment, a humanized antibody comprising a VH domain having the amino acid sequence of SEQ ID NO. 164 and a VL domain having an amino acid sequence of SEQ ID NO. 165 is used to diagnose a cancerous state, e.g. by in vivo diagnosis or for identifying a cancerous state in a tissue or cell sample taken from a patient.
In one embodiment the present invention provides a humanized antibody comprising a VH domain having the amino acid sequence of SEQ ID NO. 166 and a VL domain having an amino acid sequence of SEQ ID NO. 167. In one embodiment, a humanized antibody comprising a VH domain having the amino acid sequence of SEQ ID NO. 166 and a VL domain having an amino acid sequence of SEQ ID NO. 167 is used to treat cancer, e.g. a solid tumor cancer. In another embodiment, a humanized antibody comprising a VH domain having the amino acid sequence of SEQ ID NO. 166 and a VL domain having an amino acid sequence of SEQ ID NO. 167 is used to diagnose a cancerous state, e.g. by in vivo diagnosis or for identifying a cancerous state in a tissue or cell sample taken from a patient. I.ii STn-combotopes
As disclosed herein, the VH domain of the combotope antibody binds the carbohydrate epitope, preferably a mono-Tn, bis-Tn, mono-STn or bis-STn carbohydrate epitope.
In one preferred aspect, the VH domain of the combotope antibody binds a STn- carbohydrate epitope, such a mono-STn or bis-STn.
The VH-domain of antibody G2D11 (SEQ ID NO. 1) was by structural comparison to the VH domain of 3F1 (SEQ ID NO. 25) modified into a STn-binding VH-domain (SEQ ID NO. 28) (see example 5). Compared to G2D11 (SEQ ID NO. 1), the STn binding VH domain (SEQ ID NO. 28) has the following amino acid residue changes: I28T, A30T, P101L, de/G102, T103A and F104L.
SEQ ID NO. 28 (VH domain G2D11 mutant M2: LAL-TFT) QSDAELVKPGASVKISCKASGYTFTDHAIHWVKRKPEQGLEWIGYISPGNDDIKYNEKFKGKATL TADKSSSTAYMQLNSLTSEDSAVYFCKRSLLALDYWGQGTTLTVSS
The novel STn-combotope antibodies disclosed herein are based on these structural modification and further development, as disclosed herein, and provides specific combotope recognition, wherein the VH-chain provides recognition support for the carbohydrate part STn of the combotope on the glycoprotein antigen, while the VL- domain provides recognition support for the peptide part of the combotope on the glycoprotein antigen.
In one embodiment, the VH-domain of the combotope antibody of the present invention is a SEQ ID NO. 28-like VH-domain. In one embodiment, the amino acids of the VH- domain of the combotope of the present invention resemble the SEQ ID NO. 28-like VH- domain in their structural conformation.
As demonstrated in Example 5, amino acid residues T28, T30, L101, A102, and L103 with respect to SEQ ID NO. 28, are key amino acids for the STn-specificity. They are need for accommodation of the sialyl - i.e. for making extra space for sialyl for STn binding.
In one embodiment, the VH-domain of the combotope antibody comprises an amino acid sequence having at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 28, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 28; and comprises the amino acid residues T28, T30, H32, A33, H35, Y50, S99, L101, A102 and L1O3, with respect to SEQ ID NO. 28. Said combotope antibody will provide recognition support for mono-STn epitopes.
Hence, in one embodiment, antibodies for targeting tumor cells are provided, said antibody comprsing a VH and VL domain; wherein the VH domain of the antibody is a STn-binding domain and the amino acid sequence of the VH-domain has at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 28, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 28; wherein the amino acid sequence comprises the amino acid residues T28, T30, H32, A33, H35, Y50, S99, L101, A102 and L103, with respect to SEQ ID NO. 28; and wherein the VL domain of the antibody binds a peptide backbone associated with the STn-carbohydrate on the tumor cell. The sequence of the VH domain is preferably such that there is no binding to the peptide epitope.
In one embodiment, the VH-domain of the combotope antibody comprises an amino acid sequence having at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 28, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 28; and comprises the amino acid residues T28, T30, S52, N55, D57, L101, A102 and L103, with respect to SEQ ID NO. 28. Said combotope antibody will provide recognition support for the mono- STn epitope.
Hence, in one embodiment, antibodes for targeting tumor cells are provided, said antibody comprsing a VH and VL domain; wherein the VH domain of the antibody is a STn-binding domain and the amino acid sequence of the VH-domain has at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 28, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 28; wherein the amino acid sequence comprises the amino acid residues T28, T30, S52, N55, D57, L101, A102 and L103, with respect to SEQ ID NO. 28; and wherein the VL domain of the antibody binds a peptide backbone associated with the STn-carbohydrate on the tumor cell. The sequence of the VH domain is preferably such that there is no binding to the peptide epitope.
In one embodiment, the VH-domain of the combotope antibody comprises an amino acid sequence having at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 28, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 28; and comprises the amino acid residues T28, T30, H32, A33, H35, Y50, S52, N55, D57, S99, L101, A102 and L103, with respect to SEQ ID NO. 28. Said combotope antibody will also provide recognition support for bis-STn epitopes.
In one embodiment, the amino acid sequence of the VH-domain of the combotope antibody has at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 28, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 28; and comprises the amino acid residues T28, T30, H32, A33, H35, Y50, S52, N55, D57, S99, L101, A102 and L103, with respect to SEQ ID NO. 28.
In one embodiment, the VH domain of the combotope antibody comprises an amino acid sequence having at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 28, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 28; and in pairwise alignment with SEQ ID NO. 28 the amino acid sequence of the VH-domain comprises amino acid residues threonine (T), threonine (T), histidine (H), alanine (A), histidine (H), tyrosine (Y), serine (S), asparagine (N), aspartic acid (D), serine (S), leucine (L), alanine (A), and leucine (L) at positions corresponding to amino acid position T28, T30, H32, A33, H35, Y50, S52, N55, D57, S99, L101, A102 and L103 of SEQ ID NO. 28, respectively. The pairwise sequence alignment is performed using scoring matrix: blosum62, gap opening penalty: 10, and gap extension penalty 0.2.
In one embodiment, in pairwise alignment with SEQ ID NO. 28, the amino acid sequence of the VH-domain of the combotope antibody comprises amino acid residues threonine (T), threonine (T), histidine (H), alanine (A), histidine (H), tyrosine (Y), serine (S), asparagine (N), aspartic acid (D), serine (S), leucine (L), alanine (A), and leucine (L), at positions corresponding to amino acid position T28, T30, H32, A33, H35, Y50, S52, N55, D57, S99, L101, A102 and L103, of SEQ ID NO. 28, respectively; and the amino acid sequence of the VH domain has at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 28, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 28.
Hence, in one embodiment, antibodes for targeting tumor cells are provided, said antibody comprsing a VH and VL domain; wherein the VH domain of the antibody is a STn-binding domain and the amino acid sequence of the VH-domain has at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 28, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 28; wherein the amino acid sequence comprises the amino acid residues T28, T30, H32, A33, H35, Y50, S52, N55, D57, S99, L101, A102 and L103, with respect to SEQ ID NO. 28; and wherein the VL domain of the antibody binds a peptide backbone associated with the STn-carbohydrate on the tumor cell. The sequence of the VH domain is preferably such that there is no binding to the peptide epitope.
Combotope antibodies of the present invention as described herein, such as the STn- combotopes disclosed herein, comprise improved binding affinity to a specific antigen epitope, the combotope (compared to other epitopes on healthy or cancer cells). In some embodiment, the antibody comprises a binding affinity (e.g., kD) of between 100 nM to IpM, such as less than 100 nM, less than 10 nM, less than 1 nM, less than 100 pM, or even less than 10 pM..
In some embodiment, the antibodies of the present invention are used to treat cancer. In some instances, the cancer is lung, head and neck squamous cell, colorectal, melanoma, liver, classical Hodgkin lymphoma, kidney, gastric, cervical, merkel cell, B- cell lymphoma, or bladder cancer. In a preferred embodiment, the cancer is a solid tumor.
In another embodiment, an antibody comprising (i) a VH domain having an amino acid sequence of SEQ ID NO. 28 and (ii) a VL domain having an amino acid sequence of any one of SEQ ID NOs. 9-21 is used to diagnose a cancerous state. The antibody may be used for in vivo diagnosis or for identifying a cancerous state ex vivo in a tissue or cell sample taken from a patient.
Binding of the combotope antibody to a cancer target for diagnosis can be monitored or identified by way of known methods, such as for example labelling of the combotope antibodies and/or by use of labelled antibodies for binding of the combotope antibodies.
In yet another aspect, the present invention discloses specific antibodies.
In one embodiment, the combotope antibody of the present invention comprises
(i) a VH-domain having at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 28, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 28, and comprising the amino acid residues T28, T30, H32, A33, H35, Y50, S52, N55, D57, S99, L101, A102 and L103, with respect to SEQ ID NO. 28, and
(ii) a VL domain having an amino acid sequence selected from of any one of SEQ ID NOs. 22-24. In one embodiment the present invention provides an antibody comprising a VH domain having the amino acid sequence of SEQ ID NO. 28 and a VL domain having an amino acid sequence seleted from any one of SEQ ID NO. 22-24.
In one embodiment, the present invention provides antibodies, wherein the antibody comprises (i) a VH-domain having at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 28, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 28, and comprising the amino acid residues T28, T30, H32, A33, H35, Y50, S52, N55, D57, S99, L101, A102 and L103, with respect to SEQ ID NO. 28, and (ii) a VL domain having an amino acid sequence selected from of any one of SEQ ID NOs. 22-24; and wherein the antibody is a monoclonal antibody, a polyclonal antibody, a bi-specific antibody, a multispecific antibody, a grafted antibody, a human antibody, a humanized antibody, a synthetic antibody, a chimeric antibody, a camelized antibody, a single-chain Fvs (scFv), a single chain antibody, a Fab fragment, a F(ab')2 fragment, a Fd fragment, a Fv fragment, a single-domain antibody, an isolated complementarity determining region (CDR), a diabody, a fragment comprised of only a single monomeric variable domain, disulfide-linked Fvs (sdFv), an intrabody, an anti-idiotypic (anti-Id) antibody, or ab antigen-binding fragments thereof.
In some embodiment, an antibody comprising (i) a VH domain having an amino acid sequence of SEQ ID NO. 28 and (ii) a VL domain having an amino acid sequence of any one of SEQ ID NOs. 22-24 is used to treat cancer. In some instances, the cancer is lung, head and neck squamous cell, colorectal, melanoma, liver, classical Hodgkin lymphoma, kidney, gastric, cervical, merkel cell, B-cell lymphoma, or bladder cancer. In a preferred embodiment, the cancer is a solid tumor.
Preferably, the selected antibody of the present invention is a humanized antibody, such as mententioned above.
SEQ ID NO. 22 (VL-domain C4-STnMUCl):
IVMTQSPSSLAVSAGEKVTMSCKSSQSLLNSRTRKNYLAWYQQKPGQSPKLLIYWASTRHTGVP DRFTGSGSGTDFTLTISNVQSEDLAEYFCQQYNSYPYTFGGGTKLEIKR
SEQ ID NO. 23 (VL-domain D3-STnMUCl):
IVMTQSPSSLAVSAGEKVTMSCKSSQSLLNSRTRKNYLAWYQQKPGQSPKLLIYWASTRHTGVP DRFTGSGSGTDFTLTISNVQSEDLAEYFCQQYNSYPYTFGGGTKLEIKR
SEQ ID NO. 24 (VL-domain C7-STnMUCl):
VWTQTPLSLPVSLGDQASISCRSSQSLVHSNGNTYLHWYLQKPGQSPKLLIYKVSNRFSGVPDR FSGSGSGTDFTLKISRVEAEDLGVYFCSQSTHVPRTFGGGTKLEIKR I.iii Tn-combotopes or STn-combotopesIt may be the case that it is not known whether the combotope comprises Tn or STn as the short truncated O-glycan on the surface of a glycoprotein or that a particular glycoprotein comprisees a mixture or these glycans. In such a case, it may be beneficial to use an antibody that binds either a Tn or a STn epitope in the combotope. Thus, In a particular embodiment of the present invention, the combotope antibody comprises both a VH domain for the Tn epitope and a VH domain for the STn epitope. Such antibodies comprise a VH(Tn)-VLl arm and a VH(STn)- VL2 arm, where the VL1 and VL2 domains may be the same or different and for binding the same or different peptide epitope(s). The VH(Tn)-VLl arm and the VH(STn)-VL2 arm may be comprised as the two Fab arms in a normal antibody, as F(ab)2, as a minibody, a diabody, a triabody or be two scFv linked in one molecule, e.g. scFv-Fc . In one embodiment, the VH(Tn) domain has an amino acid sequence of SEQ NO. 1 or a functional variant thereof as defined above and the VH(STn) has an amino acid sequence of SEQ NO. 28 or a functional variant thereof as defined above.
II. Antibody library
In one aspect, the present invention provides an antibody library for in-vitro identification of a specific antibody which binds glycoproteins, such as glycoproteins on cancer tumor cells. Hence, in one embodiment, the present invention provides an antibody library for in-vitro identification of a specific antibody which binds one or more tumor cells.
As disclosed herein, tumor cells often comprise Tn or STn glycosylation epitopes on specific glycoproteins on the surface of the cancer cells. Specifically, tumor cells often comprise short truncated O-glycans on the surface, exposing for example Tn and/or STn epitopes on specific glycoproteins on the surface of the cancer cells. The antibody library of the present invention facilitates identification of antibodies with improved specificity towards tumor cells by virtue of the antibodies being specific both towards the carbohydrate epitope (Tn and/or STn) as well as towards the peptide epitope in the protein backdone associated with the carbohydrate (epitope). Specifically, such improved antibodies comprise a VH domain which efficiently binds the carbohydrate epitope of the glycoprotein, and a VL domain which efficiently binds the peptide epitope of the glycoprotein associated with the carbohydrate epitope, and the antibody is thereby specific for the combination of said carbohydrate epitope and said peptide epitope of said glycoprotein, the combined epitopes being termed a "combotope".
As mentioned previously, the VH-domain of antibody G2D11 (SEQ ID NO. 1) was by structural characterization found to particularly recognize the carbohydrate epitope of the glycoprotein epitope bis-Tn on MUC1, while it did not recognize the peptide sequence of said glycoprotein epitope (see examples 1).
Based on these structural observations, the antibody library of the present invention was conceptualized, wherein each antibody of the library comprises a specific preselected VH chain which provides recognition support for the carbohydrate part of the combotope antigen, while the VL-domain is variable, including one or more VL domains being specific for a specific peptide sequence of a particular glycoprotein, creating a library which can be screened for specific combotope antibodies, of which the VL-domain will provide recognition support to the peptide epitope within the combotope of said glycoprotein.
Based on the binding and structural data presented herein (i.e. key amino acids for Tn and STn binding, in combination with the specific G2D11 versatile VH sequence), the antibody library of the present invention can be used to identify specific glycoprotein combotope antibodies, where the VL-domain is specific for the peptide epitope, i.e. glycoprotein, of choice.
In one aspect, the invention provides an antibody library, wherein each of the antibodies in the library comprises two antibody domains: (i) a first antibody domain which binds the carbohydrate epitope of a glycoprotein on a cancer cell, i.e. Tn, bis. Tn, STn or bis. STn, and (ii) a second antibody domain selected from a repertoire of antibody domains, wherein the repertoire of antibody domains comprises one or more domains which binds a peptide epitope of said glycoprotein; wherein said library is for in-vitro identification of a specifc antibody from said library for targeting cancer cells, and wherein the specific antibody is specific for the combination of said carbohydrate epitope and said peptide epitope of said glycoprotein - i.e. the specifc antibody is a combotope antibody.
In one aspect, the invention provides an antibody library, wherein each of the antibodies in the library comprises two antibody domains: (i) a first antibody domain which binds the carbohydrate epitope of a glycoprotein of the tumor cell, and (ii) a second antibody domain selected from a repertoire of antibody domains, wherein the repertoire of antibody domains comprises one or more domains which binds a peptide epitope of said glycoprotein of said tumor cell; wherein said library is for in-vitro identification of a specific antibody from said library for targeting tumor cells, and wherein the specific antibody is specific for the combination of said carbohydrate epitope and said peptide epitope of said glycoprotein - i.e. the specifc antibody is a combotope antibody. In one embodiment, the first antibody domain is a VH domain and the second antibody domain is a VL domain. In another embodiment, both the first and second antibody are VH domains, but different from one another.
In one aspect, the present invention provides an antibody library, wherein each of the antibodies in the antibody library comprises
(i) a VH-domain which binds a carbohydrate epitope on a glycoprotein of the tumor cell, and
(ii) a VL-domain selected from a repertoire of VL-domains, wherein the repertoire of VL-domains comprises one or more VL-domains which binds a peptide epitope on said glycoprotein of the tumor cell, for in-vitro identification of a specific antibody from said library which binds a tumor cell, wherein the specific antibody is specific for the combination of said carbohydrate epitope and said peptide epitope of said glycoprotein, the combined epitope being termed a "combotope".
In one embodiment, the present invention provides an antibody library for in-vitro identification of a specific antibody which binds a tumor cell, wherein each of the antibodies in the antibody library comprises (ii) a VL-domain selected from a repertoire of VL-domains, wherein the repertoire of VL-domains comprises one or more VL-domains which binds a peptide epitope on a glycoprotein of a tumor cell, and (i) a VH-domain which binds a carbohydrate epitope on said glycoprotein of said tumor cell and which does not contribute or interfere with binding said peptide epitope of said glycoprotein of said turner cell, wherein the specific antibody is specific for the combination of said carbohydrate epitope and said peptide epitope of said glycoprotein, the combined epitope being termed a "combotope" i.e. the specifc antibody is a combotope antibody.
The VH chain of each antibody encoded by the library thereby pre-selects for a desired glycoform specificity, while the VL chain is selected from a repertoire of LV domains and will determine the peptide backbone specificity and thereby the glycoprotein specificity.
In one embodiment, the VH chain of each antibody encoded by the library pre-selects for a desired glycoform specificity, without interfering with the peptide backbone specificity, while the VL chain is selected from a repertoire of LV domains and will determine the peptide backbone specificity and thereby the glycoprotein specificity.
The structural studies disclosed herein (Example 1) provide clear evidence that the VH domain of G2D11 has no interaction with the peptide/protein carrier, and that peptide/protein interaction of the antibodies entirely comes from contribution via the VL-chain, as exemplified with antibodies obtain from the library of the present invention (e.g. antibodies Tn-MUCl and Tn-CD43; see Examples 3 and 4).
In one preferred embodiment, the the peptide epitope and carbohydrate epitope are a common continuous epitope, i.e. where the carbohydrate is in close proximity to the peptide by virtue of being chemically linked, such as by a covalent bond. In another emodiment, the peptide epitope and carbohydrate epitope may be a common discontinuous epitope, where the "associated with" still refers to the peptide epitope and carbohydrate epitope being in close proximity of one another, but by virtue of the structural configuration of the molecule facilitating the common epitope.
The antibodies herein are selected from a monoclonal antibody, a polyclonal antibody, a bi-specific antibody, a multispecific antibody, a grafted antibody, a human antibody, a humanized antibody, a synthetic antibody, a chimeric antibody, a camelized antibody, a single-chain Fvs (scFv), a single chain antibody, a Fab fragment, a F(ab')2 fragment, a Fd fragment, a Fv fragment, a single-domain antibody, an isolated complementarity determining region (CDR), a diabody, a fragment comprised of only a single monomeric variable domain, disulfide-linked Fvs (sdFv), an intrabody, an antiidiotypic (anti-Id) antibody, and ab antigen-binding fragments thereof.
In one preferred embodiment, the antibodies encoded by the antibody library of the invention are scFv, wherein the VH-domain is linked to the VL-domain.
In one preferred embodiment, the libraries disclosed herein comprise scFv antibodies, comprising a VH domain and a VL domain, where both domains are present in a single polypeptide chain. In some embodiments, the Fv polypeptide further comprises a polypeptide linker between the VH and VL domains allowing the scFv to form the desired structure for antigen binding. In one embodiment, the linker is selected from (GGGGS)n where in is 1, 2, 3, 4, 5, or 6. Many other linkers could be used as an alternative to the exemplified (GGGGS)n-linker. The skilled person would know how to select such linkers.
In one embodiment, the VH domain of each antibody in the antibody library specifically binds one or more carbohydrate epitope selected from Tn and/or STn. In one embodiment, the VH domain of each antibody encoded by the antibody library specifically binds one or more Tn moieties, such as a mono-Tn epitope or a bis-Tn epitope. In another embodiment, the VH domain of each antibody encoded by the antibody library specifically binds one or more STn-moieties, such as a mono-STn epitope or a bis-STn epitope. In yet another embodiment, the VH domain of some of the antibodies encoded by the antibody library specifically bind one or more Tn-moieties, while the VH domain of other antibodies encoded by the antibody library specifically bind one or more STn-moieties. In yet another embodiment, the antibody comprises two different VH domains, one for Tn or bis-Tn and the other for STn or bis-STn.
As disclosed above, the VH domain is specified for each antibody in the antibody library to specifically bind a specific carbohydrate epitope being characteristic of cancer cells. On the contrary, the VL domains of the antibodies in the antibody library vary from one antibody to the other, representing a repertoire of VL domains recognizing different peptides from the glycosylated protein in question, such that a repertoire of VL-domains is generated, to be screened with the intent of identifying combotope antibodies which have specificity towards glycoproteins on tumor cells by virtue of the VH domain binding a carbohydrate epitope of the glycoprotein on the tumor cell and the VL domain binding a peptide epitope on said glycoprotein on the tumor cell.
In one embodiment, the repertoire of VL-domains is generated from a naive immune repertoire, an immunized immune repertoire in a suitable animal, or a synthetically produced repertoire.
In one embodiment, the repertoire of VL-domains encoded by the antibody library is a naive immune repertoire from an animal, such as a mouse or human. Other suitable animals are pigs, rats, dogs, horses, rabbits known to the skilled artisan.
In one preferred embodiment, the antibody library is phage display library.
II. i Tn-template antibody library
Provided herein are Tn-template antibody libraries, wherein the first domain of each antibody in the library is a Tn-binding domain, while the second domain is selected from a repertoire of antibody domains comprising one or more domain binding a peptide epitope on a glycoprotein of interest associated with the Tn epitope, as disclosed above.
Provided herein are Tn-template antibody libraries, wherein the VH domain of each antibody in the library is a Tn-binding VH domain, while the VL domain is selected from a repertoire of VL domains comprising one or more VL-domains binding a peptide epitope on a glycoprotein of interest associated with the Tn epitope, as disclosed above.
In one embodiment, the present invention provides mono-Tn-template antibody libraries, wherein the VH domain of each antibody in the library is a mono-Tn-binding VH domain, while the VL domain is selected from a repertoire of VL domains comprising one or more VL-domains binding a peptide epitope on a glycoprotein of interest associated with the Tn epitope (supra').
In one embodiment, the present invention provides bis-Tn-template antibody libraries, wherein the VH domain of each antibody in the library is a bis-Tn-binding VH domain, while the VL domain is selected from a repertoire of VL domains comprising one or more VL-domains binding a peptide epitope on a glycoprotein of interest associated with the Tn epitope (supra).
In preferred embodiment, the present invention provides antibody libraries which may be used to select for antibodies which binds mono-Tn and bis-Tn epitopes, wherein the VH domain of each antibody in the library is a mono- as well as bis-Tn-binding VH domain, while the VL domain is selected from a repertoire of VL domains comprising one or more VL-domains binding a peptide epitope on glycoprotein of interest associated with the Tn epitope (supra).
In one embodiment, the VH-domain of each antibody in the Tn-template antibody library is a G2D11-Iike VH-domain (SEQ ID NO. 1). In one embodiment, the amino acids of the VH-domain of each antibody in the library resemble the G2D11-Iike VH-domain in their structural conformation.
In one embodiment, the VH-domain of the Tn-template antibody library comprises an amino acid sequence having at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 1, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 1; and comprises amino acid residues H32, A33, H35, Y50, and S99, with respect to SEQ ID NO. 1.
Hence, in one embodiment, an antibody library is provided for selecting tumor-targeting antibodies, wherein each antibody of the antibody library comprises a VH and a VL domain; wherein the VH domain of each antibody is a Tn-binding VH domain and the VH-domain comprises an amino acid sequence having at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 1, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 1, and comprises amino acid residues H32, A33, H35, Y50, and S99, with respect to SEQ ID NO. 1; and wherein the VL domain is selected from a repertoire of VL domains, wherein the repertoire comprises at least one VL domain which binds a peptide backbone epitope associated with the Tn-carbohydrate on the tumor cell. The sequence of the VH domain is preferably such that there is no binding to the peptide epitope. In one embodiment, the VH-domain of the Tn-template antibody library comprises an amino acid sequence having at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 1, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 1; and comprises amino acid residues S52, N55, and D57, with respect to SEQ ID NO. 1.
Hence, in one embodiment, an antibody library is provided for selecting tumor-targeting antibodies, wherein each antibody of the antibody library comprises a VH and a VL domain; wherein the VH domain of each antibody is a Tn-binding VH domain and the VH-domain comprises an amino acid sequence having at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 1, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 1, and comprises amino acid residues S52, N55, and D57, with respect to SEQ ID NO. 1; and wherein the VL domain is selected from a repertoire of VL domains, wherein the repertoire comprises at least one VL domain which binds a peptide backbone epitope associated with the Tn-carbohydrate on the tumor cell. The sequence of the VH domain is preferably such that there is no binding to the peptide epitope.
In one embodiment, the VH-domain of the Tn-template antibody library comprises an amino acid sequence having at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 1, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 1; and comprises amino acid residues H32, A33, H35, Y50, S52, N55, D57, and S99, with respect to SEQ ID NO. 1.
Hence, in one preferred embodiment, an antibody library is provided for selecting tumortargeting antibodies, wherein each antibody of the antibody library comprises a VH and a VL domain; wherein the VH domain of each antibody is a Tn-binding VH domain and the VH-domain comprises an amino acid sequence having at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 1, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 1, and comprises amino acid residues H32, A33, H35, Y50, S52, N55, D57, and S99, with respect to SEQ ID NO. 1; and wherein the VL domain is selected from a repertoire of VL domains, wherein the repertoire comprises at least one VL domain which binds a peptide backbone epitope associated with the Tn-carbohydrate on the tumor cell. The sequence of the VH domain is preferably such that there is no binding to the peptide epitope. In one embodiment, the VH domain of the Tn-template antibody library comprises an amino acid sequence having at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 1, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 1; and in pairwise alignment with SEQ ID NO. : 1, the amino acid sequence of the VH-domain comprises amino acid residues histidine (H), alanine (A), histidine (H), tyrosine (Y), serine (S), asparagine (N), aspartic acid (D), and serine (S) at positions corresponding to amino acid position H32, A33, H35, Y50, S52, N55, D57, and S99 of SEQ ID NO. 1, respectively. The pairwise sequence alignment is performed using scoring matrix: blosum62, gap opening penalty: 10, and gap extension penalty 0.2.
Hence, in one embodiment, an antibody library is provided for selecting tumor-targeting antibodies, wherein each antibody of the antibody library comprises a VH and VL domain; wherein the VH domain of each antibody is a Tn-binding VH domain and the amino acid sequence of the VH-domain in pairwise alignment with SEQ ID NO. 1 comprises amino acid residues histidine (H), alanine (A), histidine (H), tyrosine (Y), serine (S), asparagine (N), aspartic acid (D), and serine (S) at positions corresponding to amino acid position H32, A33, H35, Y50, S52, N55, D57, and S99 of SEQ ID NO. 1, respectively; and the amino acid sequence of the VH domain has at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 1, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 1.
In one embodiment the first antibody domain is a mono- or bis-Tn-binding VH domain, and wherein said VH-domain comprises an amino acid sequence having at least 90% sequence homology to SEQ ID NO. 1 and comprising amino acid residues H32, A33, H35, Y50, and S99 and/or amino acid residues S52, N55, and D57, with respect to SEQ ID NO. 1, preferably comprising amino acid residues H32, A33, H35, Y50, and S99 and amino acid residues S52, N55, and D57, with respect to SEQ ID NO. 1.
The Tn-template library is useful for screening for antibodies which bind combotopes comprising a Tn epitope.
II. ii STn-template antibody library
Provided herein are STn-template antibody libraries, wherein the first domain of each antibody in the library is a STn-binding domain, while the second domain is selected from a repertoire of antibody domains comprising one or more domain binding a peptide epitope on a glycoprotein of interest associated with the STn epitope, as disclosed above. Provided herein are STn-template antibody libraries, wherein the VH domain of each antibody in the library is a STn-binding VH domain, while the VL domain is selected from a repertoire of VL domains comprising one or more VL-domains binding a peptide epitope on a glycoprotein of interest associated with the STn epitope, as describes above.
In one embodiment, the present invention provides mono-STn-template antibody libraries, wherein the VH domain of each antibody in the library is a mono-STn-binding VH domain, while the VL domain is selected from a repertoire of VL domains comprising one or more VL-domains binding a peptide epitope on a glycoprotein of interest associated with the STn epitope (supra').
In one embodiment, the present invention provides bis-STn-template antibody libraries, wherein the VH domain of each antibody in the library is a bis-STn-binding VH domain, while the VL domain is selected from a repertoire of VL domains comprising one or more VL-domains binding a peptide epitope on a glycoprotein of interest associated with the STn epitope (supra).
In a preferred embodiment, the present invention provides antibody libraries which may be used to select for antibodies which binds mono-STn and bis-STn epitopes, wherein the VH domain of each antibody in the library is a mono- as well as bis-STn-binding VH domain, while the VL domain is selected from a repertoire of VL domains comprising one or more VL-domains binding a peptide epitope on a glycoprotein of interest associated with the STn epitope (supra).
As disclosed herein, the VH-domain of antibody 3F1 (SEQ ID NO. 25) efficienly binds STn. In one embodiment, the VH-domain of each antibody in the STn-template antibody library is a 3F1-Iike VH-domain (SEQ ID NO. 25). In one embodiment, the amino acids of the VH-domain of each antibody in the library resemble the 3F1-Iike VH-domain in their structural conformation.
A potential disadvantage of 3F1 is that is does not express well, however, G2D11 is very stable and easy to produce compared to 3F1. For example, G2D11 is very well expressed in Pichia pastoris, while 3F1 does not express so well. In addition, we wanted to learn the molecular basis of how to conver an anti-Tn to an anti-STn.
By sequence comparison of G2D11 and 3F1, amino acid residues were identified which might affect Tn vs STn specificity. A modified G2D11 VH domain was prepared, to simulate the 3F1 VH domain, for obtaining STn specificity (see example 5). This altered G2D11 VH domain is provided herein as SEQ ID NO. 28. Compared to G2D11 (SEQ ID NO. 1), the altered STn binding VH domain has the following amino acid residue changes: I28T, A30T, P101L, de/G102, T103A and F104L. In one embodiment, the VH- domain of each antibody in the STn-template antibody library is a SEQ ID NO. 28-like VH-domain. In one embodiment, the amino acids of the VH-domain of each antibody in the library resemble SEQ ID NO. 28 in their structural conformation.
In one embodiment, the VH-domain of the STn-template antibody library comprises an amino acid sequence having at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 28, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 28, and comprises amino acid residues T28, T30, H32, A33, H35, Y50, S99, L101, A102 and L103, with respect to SEQ ID NO. 28.
Hence, in one embodiment, an antibody library is provided for selecting tumor-targeting antibodies, wherein each antibody of the antibody library comprises a VH and a VL domain; wherein the VH domain of each antibody is a STn-binding VH domain and the VH-domain comprises an amino acid sequence having at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 28, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 28, and comprising amino acid residues T28, T30, H32, A33, H35, Y50, S99, L101, A102 and L103, with respect to SEQ ID NO. 28; and wherein the VL domain is selected from a repertoire of VL domains, wherein the repertoire comprises at least one VL domain which binds a peptide backbone epitope associated with the STn- carbohydrate on the tumor cell.
In one embodiment, the VH-domain of the STn-template antibody library comprises an amino acid sequence having at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 28, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 28, and comprises amino acid residues T28, T30, S52, N55, D57, L101, A102 and L103, with respect to SEQ ID NO. 28.
Hence, in one embodiment, an antibody library is provided for selecting tumor-targeting antibodies, wherein each antibody of the antibody library comprises a VH and a VL domain; wherein the VH domain of each antibody is a STn-binding VH domain and the VH-domain comprises an amino acid sequence having at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 28, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 28, and comprising amino acid residues T28, T30, S52, N55, D57, L101, A102 and L103, with respect to SEQ ID NO. 28; and wherein the VL domain is selected from a repertoire of VL domains, wherein the repertoire comprises at least one VL domain which binds a peptide backbone epitope associated with the STn-carbohydrate on the tumor cell.
In one embodiment, the VH-domain of the STn-template antibody library comprises an amino acid sequence having at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 28, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 28, and comprises amino acid residues T28, T30, H32, A33, H35, Y50, S52, N55, D57, S99, L101, A102 and L103, with respect to SEQ ID NO. 28.
Hence, in one embodiment, an antibody library is provided for selecting tumor-targeting antibodies, wherein each antibody of the antibody library comprises a VH and a VL domain; wherein the VH domain of each antibody is a STn-binding VH domain and the VH-domain comprises an amino acid sequence having at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 28, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 28, and comprising amino acid residues T28, T30, H32, A33, H35, Y50, S52, N55, D57, S99, L101, A102 and L103, with respect to SEQ ID NO. 28; and wherein the VL domain is selected from a repertoire of VL domains, wherein the repertoire comprises at least one VL domain which binds a peptide backbone epitope associated with the STn- carbohydrate on the tumor cell.
In one embodiment, the VH domain of the STn-template antibody library comprises an amino acid sequence having at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 28, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 28, and in pairwise alignment with SEQ ID NO. 28, the amino acid sequence of the VH-domain comprises amino acid residues Thr, Thr, His, Ala, His, Tyr, Ser, Asn, Asp, Ser, Leu, Ala and Leu at positions corresponding to amino acid positions 28, 30, 32, 33, 35, 50, 52, 55, 57, 99, 101, 102 and 103 of SEQ ID NO. 28, respectively. The pairwise sequence alignment is performed using scoring matrix: blosum62, gap opening penalty: 10, and gap extension penalty 0.2.
Hence, in one embodiment, an antibody library is provided for selecting tumor-targeting antibodies, wherein each antibody of the antibody library comprises a VH and a VL domain; wherein the VH domain of each antibody is a STn-binding VH domain and the amino acid sequence of the VH-domain in pairwise alignment with SEQ ID NO. 28 comprises amino acid residues Thr, Thr, His, Ala, His, Tyr, Ser, Asn, Asp, Ser, Leu, Ala and Leu at positions corresponding to amino acid positions 28, 30, 32, 33, 35, 50, 52, 55, 57, 99, 101, 102 and 103 of SEQ ID NO. 28, respectively; and the amino acid sequence of the VH domain has at least 60, 65, 70, 75, 80, 85, 90, 92, 94, 96, or 98% sequence homology to SEQ ID NO. 28, preferably at least 80%, more preferably at least 90%, most preferably at least 95% sequence homology to SEQ ID NO. 28.
In one embodiment, the first antibody domain is a mono- or bis-STn-binding VH domain, and wherein said VH-domain comprises an amino acid sequence having at least 90% sequence homology to SEQ ID NO. 28 and comprising amino acid residues (i), T28, T30, H32, A33, H35, Y50, S99, L101, A102 and L103, (ii), T28, T30, S52, N55, D57, L101, A102 and L103, or (iii) T28, T30, H32, A33, H35, Y50, S52, N55, D57, S99, L101, A102 and L103, with respect to SEQ ID NO. 28, preferably comprising amino acid residues T28, T30, H32, A33, H35, Y50, S52, N55, D57, S99, L101, A102 and L103, with respect to SEQ ID NO. 28.
The STn-template library is useful for screening for antibodies which bind combotopes comprising a STn epitope. The library may further be useful for screening for antibodies which bind combotopes comprising a Tn epitope, or a combination of Tn and STn epitope.
III. Nucleic acid library encoding antibodies
In one ascpect, the present invention provides a nucleic acid library encoding the antibody library disclosed herein.
All features and embodiemnts of the antibody library disclosed in section II equally applies to the nucleic acid library encoding antibodies disclosed in this section.
In one embodiment, the present invention provides a nucleic acid library encoding antibodies, wherein each of the nucleic acids in the library comprises
(i) a first nucleic acid sequence encoding an antibody domain which binds a carbohydrate epitope of a glycoprotein of the tumor cell, and
(ii) a second nucleic acid sequence selected from a repertoire of nucleic acids sequences comprising one or more nucleic acid sequences encoding an antibody domain that binds a peptide epitope of said glycoprotein of said tumor cell, for in-vitro identification of a specific antibody from said library which specifically binds a tumor cell, wherein the peptide epitope of the glycoprotein of the tumor cell is associated with the carbohydrate epitope of said glycoprotein of said tumor cell, and wherein the specific antibody is specific for the combination of the carbohydrate epitope and the peptide epitope of said glycoproteine. In one embodiment, the present invention provides a nucleic acid library encoding antibodies, wherein each of the nucleic acids in the library comprises
(i) a first nucleic acid sequence encoding a VH-domain which binds a carbohydrate epitope of a glycoprotein of the tumor cell, and
(ii) a second nucleic acid sequence selected from a repertoire of nucleic acids sequences comprising one or more nucleic acid sequences encoding a VL-domain that binds a peptide epitope of said glycoprotein of said tumor cell, for in-vitro identification of a specific antibody from said library which specifically binds a tumor cell, wherein the peptide epitope of the glycoprotein of the tumor cell is associated with the carbohydrate epitope of said glycoprotein of said tumor cell, and wherein the specific antibody is specific for the combination of the carbohydrate epitope and the peptide epitope of said glycoproteine.
Provided herein are nucleic acid libraries comprising a plurality of nucleic acid sequences, wherein each nucleic acid sequence of the plurality of nucleic acid sequences encodes an amino acid sequence forming at least a part of an antibody as described herein.
Specifically, the present invention provides a nucleic acid library encoding a plurality of antibodies, wherein each of the plurality of nucleic acid sequences encoding antibodies comprises
(i) a first nucleic acid sequence encoding a VH-domain which binds a carbohydrate epitope on a glycoprotein of the one or more tumor cells, and
(ii) a second nucleic acid sequence selected from a repertoire of nucleic acid sequences comprising one or more nucleic acid sequences encoding a VL-domain, that binds a peptide epitope of said glycoproteins of said tumor cells, for in-vitro identification of a specific antibody from said library which binds a tumor cell, wherein the peptide epitope of the glycoprotein of the tumor cells is associated with the carbohydrate epitope of said glycoprotein of said tumor cells, and wherein the specific antibody is specific for the combination of the carbohydrate epitope and the peptide epitope of the glycoprotein.
In one embodiment, the nucleic acid library comprises in the range of 108-109 nonidentical clones, such as at least 104, 105, 105, 107, 108, 109, or more non-identical nucleic acids.
In a further embodiment, the first and the second nucleic acid sequences are linked by a nucleic sequence encoding a peptide linker connecting the encoded VH sequence with the encoded VL sequence. The peptide linker is discussed above. Such linkers are generally known to the skilled artisan.
Provided herein are further vector libraries comprising a nucleic acid library as described herein. Examplary expression vectors for inserting nucleic acid libraries disclosed herein may comprise eukaryotic or prokaryotic expression vectors. Preferably, the nucleic acid library encoding antibodies are expressed using phage display technology.
Provided herein are further cell libraries comprising a nucleic acid library as described herein.
IV. Method of identifying combotope antibodies by using antibody library of the present invention
In a further aspect, the present invention provides a method for identifying an antibody for targeting a tumor cell, comprising preparing an antibody library as disclosed herein, and screening said library to identify one or more tumor targeting antibodies.
In one embodiment, the antibody library is prepared as a phage displayed library, and the screening comprises biopanning of the antibody library using specific tumor glycopeptides or the intact glycoprotein.
In a further embodiment, the method further comprises isolating the tumor targeting specific antibody, and optionally purifying the antibody. Antibody isolation and purification may be done by any common method recognized by a person skilled in the art.
In one embodiment, the process for identifying and isolating a tumor-specific antibody comprises the steps
(a) preparing an antibody library as disclosed herein, and
(b) identifying antibody candidates specific for target glycoprotein antigens from the library by binding assay(s).
In one embodiment, the identification of tumor cell specific antibody candiates comprises biopanning of the antibody library using a (specific) tumor glycopeptide or glycoprotein of said tumor cell, preferably O-glycosylated peptides or proteins, preferably where the glycosylation consists of short truncated O-glycan(s), such as a Tn-mucin or other O- glycosylated proteins having mucin-like motifs, as a purified peptide/protein or expressed on cell surfaces/tissues. The antibody library may be a phage display library, yeast display library, ribosomal display library, or similar, as recognized by a person skilled in the art.
In one preferred embodiment, the antibody library is a phage display library.
In one embodiment, the antibody library is a phage display library prepared by a method comprising the steps 1) mRNA isolation from a spleen (for the preparation of Tn and STn binding domains, the donor animal (e.g. mice) is immunized with a glycoprotein or glycopeptide comprising the short truncated O-glycans Tn or STn on the surface; for the preparation of the peptide binding domain, the the spleen is taken from a naive donor animal), 2) cDNA synthesis from said mRNA, 3a) amplification from said cDNA using a specific set of primers to obtain a first nucleic acid sequence encoding the VH domain, 3b) application from said cDNA using a mix of primers to obtain multiple nucleic acid sequences encoding the repertoire of VL domains, 4) assembly of the first nucleic acid sequence endocing the VH-domain and a second nucleic acid sequence from the multiple nucleic acid sequences encoding the VL-domain repertoire, to form a joint contruct, 5) insertion of the construct into a phagemid vector, 6) insertion of the phagemid vector comprising the construct into E. coli to produce a bacterial library, 7) using the bacterial library for infection of phages to produce the phage display library.
As a non-limiting example, such phage display library may be prepared as illustrated in Figure 2, comprising 1) mRNA isolation from mouse spleen. 2) cDNA synthesis with reverse transcriptase using random hexamers. 3) PCR amplification from cDNA template to obtain the VH domain using a specific set primers, and the repertoire of VL domains using a mix of VL primers. 4) PCR assembly of VL-domain repertoire and specific VH- domain using 5' phosphorylated outer primers. 5) Rolling circle amplification where phosphorylated scFv genes are ligated into circular DNA, dsDNA is denatured, random hexamers are annealed and Phi29 polymerase amplifies the circular fragments into long linear concatemers. 6) Amplified extended scFv genes are digested by sfil restriction enzyme and ligated to Sfil and rSAP treated phagemid vector pAKlOO. 7) Pool of phagemids containing scFv genes are electroporated to TGI E. coli cells. 8) Growth of Bacterial Library containing the different phagemids and infection with helper phage VCSM13 to produce complete phages displaying scFvs on their pill coat protein.
Selections from the antibody library, such as from the phage display library, may be performed by several rounds of interogation (panning) with immobilized biotinylated target antigen (eg. bisTn-MUCl glycoprotein) on streptavidin-coated magnetic beads. After each round of selection, phages are eluted, amplified and precipitated. Removing extraneous phage antibodies by absorption against non-targets (negative binders), naked beads, plastics, proteins, peptides or normal human cells may also be performed as needed. Sequencing (NGS) of enriched phages after each round of panning provides a fingerprint of VL-domain antibody sequences corresponding to target antigen structure and peptide sequence. Polyclonal phage ELISA may be used to confirm enrichments for target binder (e.g bis Tn-MUCl target protein/peptide). Phage pools may then be converted to soluble scFvs and expressed as individual scFvs. Expression of the scFvs in the supernatant may be assessed with dot blot analysis.
Screening for tumor-specific scFv clones in said phage library may be done using a ELISA binding assay, glycoprotein/peptide microarray and biolayer interferometry (OCTET) against target protein/peptide and control proteins/peptides (non targets), provided scFv antibodies targeting the selected glycoprotein antigen with high specificity and affinity.
Binding (FACS) of selected scFv to tumor cells expressing the target glycoprotein antigen, such as breast adenocarcinoma cell lines MCF7, MDA-MD-231 COSMC KO may further be used to confirm tumor specificity.
In one embodiment, the present invention provides a method as disclosed herein, wherein the antibody library is a phage displayed library, and the screening comprises biopanning of the antibody library using a specific tumor glycopeptide, such as Tn-MUCl, Tn-CD43, Tn-MUC4, Tn-MUC16, Tn-MUC13, etc. Several different Tn-tumor target are - as a non-limiting example - disclosed in the review by Kudelka et al 2015.
V. Epitope targets
The present invention provides an antibody library and method for identifying specific antibodies against combined glycoside-peptide epitopes (combotobes) on specific glycoproteins, such as Tn, bis-Tn, STn or bis-STn epitopes on specific cancer cells.
Provided herein are glycoside-peptide epitope binding antibodies which may have therapeutic effects due to their ability to specifically bind to specific glycoproteins on for example cancer cells. Preferably, the antibody library provided herein facilitates identification of an antibody that may be used to identify (diagnose) or treat a disease or disorder, such as cancer.
Provided herein are methods for treatment of proliferative disorders. Further provided herein are methods for treatment of a proliferative disorder, wherein the proliferative disorder is cancer, comprising identifying and isolating anti-cancer cell antibodies by screening an antibody library of the present invention for antibodies having high specificity for said cancer cells, and administering to a subject diagnosed with said cancer disease the antibody identified as described herein. A particular method of treatment involves the use of the combotope antibodies of the present invention in loading natural killer (NK) cells with the specific antibodies for targeting the NK-cells to the target cells, e.g. cancer cells. The NK-cells may be harvested from the patient to be treated prior to the loading with the antibodies or provided as donor NK-cells. The loading may be in the form of the antibody or antibodies per se or as (a) nucleotide sequence(s) encoding the specific antibody/antibodies. In another method, the specific antibodies are linked to a cell toxin or a non-toxic precursor thereof or a similar cytotoxic effector molecule of cell death. The skilled artisar would readily know which effector molecules could be useful. Further provided herein are methods for treatment of a proliferative disorder wherein the cancer is selected from lung, head and neck squamous cell, colorectal, melanoma, liver, classical Hodgkin lymphoma, kidney, gastric, cervical, merkel cell, B-cell lymphoma, and bladder cancer. In a preferred embodiment, the cancer is a solid tumor.
Aspects of the invention include administering any one of the specific antibodies identified as described herein to a subject identified as having aberrant/truncated O- glycosylation (e.g. trucated O-glycosylation of MUC1 protein) as compared to a reference level, (e.g. level in a non-cancerous cell). In one embodiment, the truncated O- glysylation is selected from Tn and STn antigens, such as Tn-MUCl.
In one embodiment, the present invention provides combotope antibodies as disclosed herein for use in treatment of a disease associated with aberrant/truncated O- glycosylation. In one embodiment, the present invention provides combotope antibodies as disclosed herein for use in treatment of a disease associated with Tn and/or STn antigens. In one preferred embodiment, the present invention provide combotope antibodies as disclosed herein for use in treatment of cancer.
In another aspect, the disclosure features methods that include administering any one of the specific antibodies identified as described herein, or a composition comprising such antibody, e.g. a cell composition, antibody-drug conjugate, or antibodyradioisotope conjugate) to a subject in need thereof, said subject having, or identified or diagnosed as having a cancer characterized by hypoglycosylation of peptide epitopes in the cancer cells (e.g., pancreatic cancer, epithelial cancer, breast cancer, colon cancer, lung cancer, ovarian cancer, or epithelial adenocarcinoma).
Other embodiments include using a specific antibody identified as described herein in testing for the presence of cancer in a subject, for example as a or part of a test kit In some embodiments, the libraries of the present invention comprise antibodies that are adapted to the species of an intended therapeutic target. Generally, these methods include "mammalization". In some instances, the mammal is mouse, rat, equine, sheep, cow, primate (e.g., chimpanzee, baboon, gorilla, orangutan, monkey), dog, cat, pig, donkey, rabbit, and human. Preferably, the antibodies are intended for human therapeutic targets, and therefore humanized.
VI. Use of the identified antibodies
Tumor-specific antibodies are used in immuno-oncology to target cancer cells and activate the immune system to attack these cells. They can work by directly binding to cancer cells and triggering an immune response, or by targeting molecules on cancer cells that suppress the immune response. This can lead to increased tumor cell death and/or slower tumor growth. Tumor-specific mAbs are often used in combination with other immune-based therapies, such as immune checkpoint inhibitors or CAR-T cell therapy, to enhance the anti-tumor immune response. Tumor-specific monoclonal antibodies are also used in antibody-drug conjugates (ADCs) to deliver a cytotoxic drug directly to cancer cells. The mAb in the ADC is designed to recognize and bind to a specific protein on the surface of cancer cells, and once bound, the cytotoxic drug is released to kill the cancer cell. The advantage of using an ADC is that it can selectively deliver the drug to cancer cells, minimizing the damage to healthy cells. Some examples of ADCs that use tumor-specific mAbs include trastuzumab emtansine (T-DM1) for HER2-positive breast cancer and inotuzumab ozogamicin for acute lymphoblastic leukemia.
In one embodiment, the antibodies of the present invention - i.e. antibodies identified using the antibody library of the present invention - are used to target cancer cells, such as to activate the immune system to attack the cancer cells.
Hence, in one aspect, the present invention provides an antibody as disclosed herein for use in treatment and/or prevention of cancer. In some instances, the cancer is lung, head and neck squamous cell, colorectal, melanoma, liver, classical Hodgkin lymphoma, kidney, gastric, cervical, merkel cell, B-cell lymphoma, or bladder cancer. In a preferred embodiment, the cancer is a solid tumor.
In one embodiment, the antibodies of the present invention are used in combination with other immune-based therapies, such as immune checkpoint inhibitors and/or CAR- T cell therapy, to enhance the anti-tumor immune response. In one embodiment, the antibodies of the present invention is used in antibody-drug conjugates (ADCs), such as to deliver a cytotoxic drug directly to cancer cells.
In another aspect, the present invention provides a method of treating a cancer comprising administering a formulation comprising at least one specific antibody as disclosed herein to a patient in need thereof.
In one embodiment, the the antibody administered is conjugated to a cytotoxic moiety or loaded into a NK-cell, such as the patients own NK-cells, for being presented on the surface thereof.
Specific antibodies for use in treatment of cancer may be selected from a list of antibodies, wherein all antibodies comprise
(i) a VH-domain which binds a Tn- and/or STn-carbohydrate epitope of a glycoprotein of the tumor cell, and
(ii) a VL-domain which binds a peptide epitope of said glycoprotein of said tumor cell, and wherein the antibody is specific for a combination of said carbohydrate epitope and said peptide epitope of said glycoprotein, but wherein the list of antibodies does not 5E5, 5F7, 2D9.
In one embodiment, the antibody administered to the patient in treatment of cancer comprises
(i) a VH domain comprising an amino acid sequence having at least 90% sequence homology to SEQ ID NO. 1, and amino acid residues H32, A33, H35, Y50, and S99 and/or S52, N55, and D57, with respect to SEQ ID NO. 1, and (ii) a VL domain comprising an amino acid sequence selected from SEQ ID NO. 9-21.
In one embpodiment, the antibody administered to the patient in treatment of cancer comprises
(I) a VH domain comprising an amino acid sequence having at least 90% sequence homology to SEQ ID NO. 28, and comprising amino acid residues (i), T28, T30, H32, A33, H35, Y50, S99, L101, A102 and L103, (ii), T28, T30, S52, N55, D57, L101, A102 and L103, or (iii) T28, T30, H32, A33, H35, Y50, S52, N55, D57, S99, L101, A102 and L103, with respect to SEQ ID NO. 28, and (II) a VL domain comprising an amino acid sequence selected from SEQ ID No. 22-23.
In one embodiment, the antibody administered to the patient comprises (i) a first VH domain comprising an amino acid sequence having at least 90% sequence homology to SEQ ID NO. 1, and amino acid residues H32, A33, H35, Y50, S52, N55, D57, and S99, with respect to SEQ ID NO. 1, and (ii) a first VL domain comprising an amino acid sequence selected from SEQ ID NO. 9-21; or
(I) a second VH domain comprising an amino acid sequence having at least 90% sequence homology to SEQ ID NO. 28, and comprising amino acid residues T28, T30, H32, A33, H35, Y50, S52, N55, D57, S99, L101, A102 and L103, with respect to SEQ ID NO. 28, and (II) a second VL domain comprising an amino acid sequence selected from SEQ ID No. 22-23.
In another aspect, the present invention concerns diagnostics. Specifically, a Tn- or STn- binding monoclonal antibody as disclosed herein can be utilized as a diagnostic tool for cancer in a subject by targeting the Tn/STn combotope found on the surface of many cancer cells but rarely present in normal cells. The subject may be a human or an animal. The process begins with the administration of the Tn- or STn-binding mAb, which has been designed to specifically recognize and bind to the Tn- or STn- antigen. Once administered, the mAb circulates through the body and binds to the Tn- or STn- antigens expressed on the surface of cancer cells. This binding can be detected and visualized using various imaging techniques, such as PET, MRI, or fluorescence imaging, depending on the label attached to the mAb. The presence and distribution of the Tn/STn-antigen- mAb complexes in the body can then be analyzed to determine the presence, extent, and possibly the type of cancer. This method offers a targeted approach to cancer diagnosis, potentially allowing for earlier detection and a more precise understanding of the cancer's location and spread, which is crucial for effective treatment planning. Combotope antibodies of the present invention may be used in such diagnostics approach. In a another approach to diagnosing cancer in a patient, a tissue sample is collected from the patient suspected of suffering from a cancerous state, for example in form of a biopsy from the suspected cancerous tissue or by removal of whole or parts of the cancerous tissue for subsequent diagnosis ex vivo by binding one or more combotobe antibodies of the present invention to the tissue sample followed by identification of specific binding of the antibody by methods commonly known to the skilled artisan working in the field of identifying tissue antigen targets by way of immune detection. Such diagnostic methods are generally known and performed on a daily basis in hospitals around the world.
VII. Humanization of antibodies As mentioned previously, the antibodies of the present invention are preferably humanized. This may be done is several different ways as acknowledged by a person skilled in the art.
A non-limiting example of such humanization of antibodies comprises the following steps:
1) Identification of the mouse monoclonal antibody: The first step in humanizing a mouse monoclonal antibody is to identify an antibody with the desired specificity and affinity. This is typically done by screening a large library of mouse monoclonal antibodies using techniques such as ELISA or flow cytometry.
2) Analysis of the antibody structure: Once a mouse monoclonal antibody with the desired specificity and affinity is identified, its structure is analyzed to identify the regions responsible for its antigen-binding properties. These regions are typically located in the variable regions of the antibody, which are highly diverse and are responsible for recognizing and binding to specific antigens.
3) Selection of a human antibody framework: A human antibody framework is selected based on its structural similarity to the mouse antibody framework. This is important because it ensures that the humanized antibody retains the overall structure and stability of the original antibody.
4) Replacement of the antigen-binding regions: The mouse-derived antigenbinding regions, also known as complementarity-determining regions (CDRs), are replaced with human-derived CDRs while retaining the overall structure of the antibody. This is done using genetic engineering techniques such as PCR, cloning, and site-directed mutagenesis. Often a minor number of amino acids need to be replaced. Conservative substitions may be made without changing the properties of the antibody. The skilled artisan knows have to select conservative amino acids for replacement purposes in the antibody, for example to humanize the antibody or to facilitate the synthesis and production of the antibody. The substitution(s) are made by changing the nucleotide code(s) in the nucleotide sequence encoding the domains of the antibody.
5) Testing of the humanized antibody: Once the humanized antibody is produced, it is tested for its specificity, affinity, and functionality. This is typically done using techniques such as ELISA, flow cytometry, and Western blotting. The humanized antibody is also tested for its immunogenicity, which is its ability to trigger an immune response in humans. If the humanized antibody is found to be safe and effective, it can be further developed for use in human therapies. In case of immunogenicity of a certain promising combotope antibody, some of the amino acids may be substituted by concervative counterparts, however securing the the specificity and efficaty remains unchanged or even improved.
VIII. Method of identifying glycopeptide targets
In a further aspect, the present invention provides a method for identifying a glycopeptide target, said target comprising a Tn and/or STn epitope, such as a glycopeptide target of a cancer cell. The library of the presnt invention is used for the identification of such glycopeptide targets by for example immunoprecipitation and mass spectrometry: This approach involves incubating the phage display antibody library with a cell lysate or tissue sample and allowing the antibody to bind to its target protein. The antibody-protein complex is then isolated by immunoprecipitation and subjected to mass spectrometry analysis to identify the protein. Another example is Protein microarray: Protein microarrays are arrays of immobilized antibodies that can be used to identify protein targets of antibodies. By incubating the phage display antibody library array with a cell lysate or tissue sample, and detecting binding, it is possible to identify the target protein.
Hence, in one embodiment, the present invention discloses a method for identifying a glycopeptide target, said target comprising a Tn and/or STn epitope and a peptide target, such as a glycopeptide target on a cancer cell, said method comprising the steps of i) preparing an antibody library as disclosed herein, and ii) incubating said antibody library with a sample comprising the glycopeptide target, iii) analyzing one or more antibody-peptide complexes obtained from step ii) to identify the amino acid sequence of the peptide epitope of said glycopeptide target.
In a preferred embodiment, the sample is a cell or tissue sample, such as lysed cells or tissues. Glycoproteins may be isolated from the lysed celles and used for the screening. The antibody library is preferably prepared as a phage display library, and the glycopeptides targets may be identified by analysis of the antibody-glycopeptide complexes using mass spectrometry analysis, or other similar method as recognized by a person skilled in the art.
IX. Determining VH- and VL-domain specificity
As disclosed herein, the VH-domain of the combotope antibody of the present invention binds a carbohydrate epitope (Tn and/or STn) of a glycoprotein of the cancer cell, while the VL-domain of the combotope antibody binds a peptide epitope of the same glycoprotein associated with the carbohydrate epitope.
By structural characterization, such as using X-ray crystallography (as described herein in example 1), a person skilled in the art is able to identity whether a VH-domain may be characterized as a carbohydrate epitope binding VH domain, and further whether a VL-domain may be characterized as a peptide epitope binding VL domain.
Another way of identifying a VH domain is by sequence analysis (as described herein in example 3.3). Identification of whether a sequence is a VH domain may be done by alignment of the sequence in question with (a) G2D11 VH domain SEQ ID NO.: 1, to check whether the amino acid sequence in question comprises amino acid residues positions corresponding to H32, A33, H35, Y50, S52, N55, D57, and S99, with respect to SEQ ID NO.: 1 - i.e. whether the sequence in question have the key residues needed for the VH domain functionality of Tn binding; or with (b) SEQ ID NO. 28 to check whether the amino acid sequence in question comprises amino acid residues positions corresponding to T28, T30, H32, A33, H35, Y50, S52, N55, D57, S99, L101, A102 and L103, with respect to SEQ ID NO. 28 - i.e. whether the sequence in question have the key residues needed for the VH domain functionality of STn binding. At least three amino acids corresponding to the above mentioned amino acid residues need to be present in the VH domain for it to function as a VH domain binding Tn and/or STn.
For the VH-domain to be versatile, it is important that the VH-domain does not contribute to or interfere with binding any peptide epitopes on the glycoprotein, which means that binding of the VH-domain to the carbohydrate epitope is not influenced, interfered or affected by any peptide epitopes on the glycoprotein. Only in this way, the VH-domain can be freely be combined with any VL-domain of choise for a "clean" binding without any disturbing binding between the VH-domain of the antibody and a peptide epitope, which unwanted binding may distort the result - in e.g. diagnosis or specific drug delivery, targeting a specific tumor glycoform (Tn or STn) on a given protein.
For the VL-domain, this is discovered as disclosed herein, and will be unique to each target. The identified VL domains have been (1) evaluated based on specificity and (2) correlated with other VL-domain sequences to identify common traits in the CDRs. X- ray may then confirm these traits.
EXAMPLES
The following examples are set forth to illustrate more clearly the principle and practice of embodiments disclosed herein to those skilled in the art and are not to be construed as limiting the scope of any claimed embodiments.
All chemicals were supplied by Merck, Germany unless otherwise stated. All buffers, media, were dissolved in milli Q water (MQ) and autoclaved unless otherwise stated.
Peptides
The peptides that were used in phage display selection, ELISA and Bio-layer interferometry (BLI) are shown in table 1. CD43 and IgA have been previously synthetized in house by solid phase peptide synthesis (SPPS) as described in Persson et al 2016. MUC1 peptides were either synthesized or purchased by Biosyntan, Germany.
Btn = biotin; Ahx=aminohexyl; OEG=Oligo-ethylenglycol
Cell lines
All cell lines were maintained at 37oC in a 5% CO2 humidified incubator. MCF7, MDA- MD-231 WT and COMSC KO were maintained in DMEM+GlutaMax (Gibco, 32430-027) supplemented with 10% FBS (FisherScientific, 11550356), 1% penicillin-streptomycin (FischerScientific, 15140122) and 1 mM sodium pyruvate (Gibco, 11360). MDA-MB-231 WT and COSMC KO cells were kindly provided by Ulrich auf dem Keller. Jurkat cells were maintained in RPMI (Life Technologies, 32404014) supplemented with 10% FBS, 1% penicillin-streptomycin and 2 mM L-glutamine (Sigma, G7513). HEK293 cells were maintained in Freestyle media (Thermo Scientific, 15285885).
Broth media, Plasmids and E. coll strains
XLl-Blue electrocompetent cells were supplied by Agilent (Agilent, 200228). TGI for phage display were kindly provided by Peter Kristensen from Aalborg University. Vectors pAKlOO phagemid and pJB33 expression vector were kindly provided by Plunthum from University of Zurich, both with chloramphenicol antibiotic resistance. E. coli TGI and XL1- blue electrocompetent cells were cultures in 2xYT broth media. Liquid media was supplemented with 25 pg/mL chloramphenicol and 2% glucose unless otherwise stated.
G2D11 VH chain and mutant in pTwist vector by Twist Biosciences.
Softwares
GraphPad prism 9 was used for graph design and Biorender for image design. CLC Main workbench 8.0 software was used for sequence alignment.
Example 1: Characterization of G2D11
G2D11 is a mouse derived anti-Tn-scFv mAb. ScFv consists of VH domain SEQ ID NO 1 and LV domain SEQ ID NO. 2 joined by pepide linker (GGGGS)4.
1.1 Structural characterization
ScFv G2D11 crystals were prepared by the sitting drop technique and by using appropriate precipitant solutions. The resulting crystals were used to solve the structure at a resolution of 1.9 A and interpret the density map (Figure 3). Despite two molecules were present in the asymmetric unit and that contacted weakly between each other, analytical ultracentrifugation showed that this monomeric form behaved as a monomer either in the absence or presence of the bis-Tn-MUCl peptide APGS*T*AP where * denotes a GalNAc moiety (SEQ ID NO.: 55). The glycopeptide laid within a surface groove formed by the light (L) and heavy (H) chains (hereafter VL and VH, respectively), and in particular the two GalNAc moieties were recognized by residues from the three hypervariable regions of the VH (Figure 3).
With the exception of the OH6, all Ser-bound GalNAc hydroxyl groups were engaged in hydrogen bonds. In detail, hydroxyl group OH3 interacted with the NH group of Ala33H and OH4 with the side chains of His32H and Ser99H. The endocyclic oxygen of the sugar was engaged in hydrogen bonding with Ser99H. The carbonyl group of GalNAc was involved in a hydrogen bond with the side chain of His35H and the methyl group was engaged in a CH-n stacking interaction with His50H.
These interactions for G2D11 were found conserved when compared to a previous structure solved using the scFv-5E5 complexed to a mono-Tn-MUCl peptide (APGST*AP) (Macias-Leon et al 2020). In addition for 5E5, a further CH-it interaction between PhelO2H and the Thr5 methyl group was visualized (Macias-Leon et al 2020). Phel02 helps direct VH into one of the two GalNAcs. The implications of this Phel02 is that 5E5 has a less flexible approach to Tn, it is restricted to Tn-Thr, and can only bind in a certain way. This further suggests that 5E5 prefers monoTn instead of bisTn.
Strikingly, the Thr-bound GalNAc was also intimately recognized by the scFv-G2Dll. In this case, Ser52H was engaged in hydrogen bond interactions with the carbonyl group, OH3 and OH4, while Asn55 and Asp57 side chains interacted with OH4. G2D11 is more open and can easier tolerate binding any combination TnThr/TnSer, TnSer/TnThr, TnSer/TnSer, TnThr/TnThr.
The recognition by G2D11 for GalNac binding sites was supported by the absence of binding of a triple mutant (H32AH-H35AH-S52AH) or the double mutant (S99AH-S52AH ) towards bis-Tn-peptides (data not shown).
Interestingly, scFv-G2Dll did not recognize the peptide sequence though the Alai and Pro2 were surrounded by aromatic residues of the VL (Figure 3). On the contrary, the scFv-5E5 VL recognized the peptide by a hydrogen bond between Tyr98L and Pro7 backbones and a CH-it interaction between TyrlOOL and Pro7 (Macias-Leon et al 2020).
1.2 Sequence alignment
G2D11 VH-domain was aligned with other known anti Tn-antibody VH-domains using CLUSTALW (using standard settings for multiple alignment parameters - i.e. scoring matrix: blosum62, gap opening penalty: 10, and gap extension penalty 0.2). Based on this alignment, conserved amino acid residues in the CDR1, CDR2 and CDR3 regions relevant for the functionality of the VH-chain (i.e. binding the GalNac) were identified, as illustrated in Figure 4 by the arrows.
Specifically, with reference to G2D11 (SEQ ID NO. 1), amino acid residues H32, A33, and H35 in CDR1, amino acid residues Y50, S52, N55, and D57 in the CDR2, and also amino acid residue S99 in the CDR3 should preferably be conserved for the VH domain.
Example 2: Tn-template phage display libraries conceptualization Based on structural the VH-domain X-ray data and the observation that bisTn-binding was due to the VH-domain (as disclosed in Example 1), an antibody library was conceptualized, wherein each antibody comprised the VH chain of the previously identified scFv G2D11, providing recognition support for the glycoside part of the antigen, while the VL-domain was variable originating from naive mice, creating a scFv phage display library, termed as Tn-template library, which can be screened for a specific scFv, of which the VL-domain would provide recognition support to the underlying peptide antigen within the combotope.
2.1 Phage display library construction
Wild type BALB/c mice were euthanized, the spleens were removed and directly stored in -80oC in RNAIater RNA stabilization reagent until use. The RNA was isolated from the spleen using a gentleMACS Dissociator and miRNeasy kit (Qiagen) according to the manufacture's instruction. 1 pg of RNA was used for cDNA synthesis with random hexamer primers (FisherScientific, 10609275) and Superscript IV Reverse Transcriptase (Invitrogen, 18090010). The constant VH gene as well as the VL antibody specific genes were amplified by PCR using Q5 Hot Start High-Fidelity DNA Polymerase (NEB M0494S). The primers that were used are found in the sequence listing (SEQ ID NOs. 31-32). VL and VH genes were gel exctracted and assembled with 5' phosphorylated outer primers to allow the rolling circle amplification in the next step. RCA improves restriction enzyme (Sfil) cutting of the scFv genes. The assembled scFv fragments were sub-cloned in the Sfil-digested phagemid vector pAKlOO using Electroligase (NEB M0369) for 16 h at 16oC/25oC. The phagemid pool with a variety of scFv fragments was electroporated in XLl-Blue electrocompetent cells (Agilent, 200228). Cells were recoved in SOC media, incubated for 1 h at 37oC at 220 rpm, plated on selective media agar plates and incubated overnight at 30oC. Colonies were scraped off with cold 2xYT, supplemented with 25% v/v glycerol and stored at -80oC. For phage rescue, cells were infected with VCSM13 helper phage yielding 1013 phages/mL that were used in the biopanning. The primers that were used are found in the sequence listing (SEQ ID NOs. 33-49).
2.2 Solid phase scFv antibody selection, phage rescue and production
Three rounds of selection were performed using streptavidin M-280 Dynabeads (Invitrogen 11205D) and all incubations were performed at RT on rotation unless otherwise stated.
Beads were blocked with 5% BSA in PBST (PBS with 0.05% Tween 20) for 1 h. After 3 washes with PBST, 100 nM of biotinylated peptides (50 nM in the last round) diluted in 3% BSA in PBST were coupled with the beads for 2h. Phage library (1012 phages/mL) were pre-selected against naked beads for Ih and then transferred for positive selection against the target antigen for 2 h. Unbound phages were removed by washing 3 times with 3% BSA in PBST, 3 times with PBST and 3 times with PBS. Bound phages were eluted with 1 mg/mL of freshly prepared trypsin solution and were allowed to infect exponentially growing E. coli TGI in for 30 min at 37oC. Infected bacteria were spread on agar plates and incubated overnight at 30oC. Colonies were scraped with medium, homogenized and 1: 1000 of the homogenous mixture was inoculated in liquid media and were grown at 37oC at 220 rpm until they reach 00600=0.4-0.5. Subsequently, they were infected with VCSM13 helper phage (109 phages/mL) for 30 min at 37oC. Bacteria were spun and pellet was resuspended in liquid media without glucose but supplemented with antibiotics and isopropyl B-D-l-thiogalactopyranoside (IPTG in 1: 1000 dilution) to induce phage production. Overnight cultures were centrifuged to remove bacteria pellet and phages were precipitated form the supernatant by adding ice cold PEG/NaCI (20% w/v PEG6000, 2.5 M NaCI) in 1:4 ratio. After 1 h incubation on ice, precipitated phages were spun at 10,800xg for 30 min followed by a centrifugation at 5,000xg for 5 min. phage pellet was resuspended in 1 mL of cold PBS and was further centrifuged at 13,000xg for 10 min to remove any cell debris. Concentration was measured spectrophotometrically at 269/320 nm according to the following equation: Virions/ml = (A269-A320)x6xl016/number of bases per virion. The precipitated phages were used in the subsequent rounds of selections.
For the selections with the STn library, peptides were immobilized on NHS beads (Fisher Scientific, 88827). First beads were washed once with ice cold 1 M hydrochloric acid (HCI). 100 nM of peptides were diluted in print buffer and incubated with the beads for 2 h. two washed with 0.1 M glycine pH= 2 followed and then blocking of 1 h with 3 M ethanolamine for 1 h. Phage library incubation, washes and trypsin elution followed as described previously.
2.3 Subcloning, expression and scFv screening
After three rounds of selection, polyclonal phagemids with the different scFv fragments were purified with the GeneJet Plamsid Miniprep Kit (Thermo Fischer, K0503) according to the protocol, digested with Sfil restriction enzyme for 20 min at 50oC and ligated in the pJB33 expression vector using T4 electroligase for 1 h at 65oC. The pool of the different constructs was electroporated in XLl-Blue electrocompetent cells, cells were recovered in SOC media, incubated for 1 h at 220 rpm at 37oC and then cells were spread on agar plates and incubated overnight at 37oC. 62 individual colonies were picked and incubated overnight at 37oC in 96 U-bottom well plates. Overnight cultures were inoculated in fresh media without glucose and were incubated for 4 h at 37oC at 220 rpm before IPTG induction (0.5 mM final concentration) and overnight cultivation at 30oC, at 800 rpm in a humidified incubator. Cells were pelleted by centrifugation at 3,000xg for 10 min and the supernatant was used for positive hint binding with ELISA. For clone sequencing analysis, plasmid DNA was purified from each individual clone and sent for Sanger sequencing in Macrogen using M13R custom designed primer. CLC Main workbench 8.0 software was used for sequence alignment.
2.4 Soluble scFv production and purification
All periplasmic protein extraction steps were performed on ice for 1 h and all the centrifugations at 4oC. Selected clones were grown from the glycerol stocks overnight at 37oC, at 220 rpm, in 2xYT supplemented with 2% glucose and 25 pg/ml chloramphenicol. Overnight cultures were diluted in fresh media and cultured until exponential phase before induction with ImM IPTG and overnight incubation at 20oC, at 220 rpm. Bacteria were harvested at 6,000xg for 10 min and pellet was resuspended in ice-cold 100 mM Tris, 20% w/v sucrose solution with EDTA-free protease inhibitor cocktail (ThemroFischer, A32965), pH 8. After centrifugation at 8,000xg for 10 min, pellet was resuspended in ice-cold 5 mM MgSO4 in MQ solution. Pellet was centrifuged at 8,000xg for 10 min and the 2 fractions were pooled together and centrifuged at 12,000xg for 60 min to remove any cell debris.
Pooled fractions with the soluble scFv were filtered with 0.45 pm filter membrane, the supernatant was mixed with 4x equilibration buffer (100 mM Tris, 1.2 M NaCI, pH 8) in 3: 1 ratio (v/v) and were incubated overnight with nickel-nitrilotriacetic acid (Ni-NTA) agarose beads (Qiagen, 30210) at 4oC on rotation. Beads with the captured scFv were pelleted by centrifugation at l,000xg for 2 min and were loaded on a pre-equilibrated affinity resin column (Thermo Scientific, 29920) with 10 column volumes (CV) of IX equilibration buffer. Unbound proteins were washed away with 10 CV wash buffer (lx equilibration buffer with 10 mM imidazole, pH 8) and bound scFv were eluted with 0.2 CV of elution buffer (lx equilibration buffer with 250 mM imidazole, pH 8). The elution step was repeated 2 more times. Eluted scFv antibodies were desalted followed by buffer exchange in PBS with Zeba spin desalting columns (Fisher Scientific, 89892) according to the manufacturer. Protein quantification was performed with BCA Protein assay kit (ThermoFischer, 23225) according to the protocol and purity was evaluated with SDS- PAGE.
2.5 Enzyme-linked immunosorbent assay (ELISA)
All steps were performed at RT shaking unless otherwise stated. All washes between steps were done with PBST (PBS with 0.05% Tween20). 96-well Maxisorp plates (ThermoScientific, 10394751) were used for antigen immobilization in coating buffer (0.015M Na2CO3, 0.035 M NaHCO3, pH 9.6). Streptavidin (NEB N7021S) was coated overnight at 4oC at a concentration 4 times less of the antigen concentration. PLIP (0.5M NaCI, 0.003M KCI, 0.0015M KH2PO4, 0.0065 M Na2HPO4.2H2O, 1% w/v BSA, 1% Tween20) was used as blocking buffer for 1 h shaking. 3,3', 5,5; -tetramethylbenzidine (TMB, Fisher Scientific 12617087) chromogen was used as a substrate for positive signal detection. The reaction stopped with 0.5 M H2SO4 and absorbance was measured at 450 nm on a VICTOR Nivo plate reader.
For the polyclonal phage ELISA, plates were incubated with polyclonal phages in serial dilutions in blocking buffer for 2h followed by incubation with secondary antibody incubation for 1 h. Bound phages were detected with mouse monoclonal anti M13- HRP antibody (Nordic Biosite 58-11973-MM05T-H-100) at 1: 10000 dilution.
For monoclonal scFv ELISA, antigen concentration was 50 nM. 50 uL of supernatant from the overnight culture of each clone was added per well. For antibody titration ELISA, antigen coating was at fixed concentration of 330 nM peptides and scFv were titrated 5- fold starting from 300 nM. Bound scFv were detected with a mouse monoclonal anti-His HRP (C-term) (Invitrogen 46-0707) at 1:2000 dilution.
2.6 Cell-binding assays
Cells were washed twice in FACS buffer (DPBS (Sigma-Aldrich, D8537) with 0.1% w/v BSA) before treatment with 100 mU/mL Clostridium perfringens neuraminidase (Sigma, N5631) for 30 min at 37oC. After 2 washes, cells were resuspended in 100 pL of FACS buffer and transferred in 96 U bottom well plate. Cells were stained for 2.5 pg/mL PNA (B-1075), 0.4 pg/mL WA (B-1235), 1 pg/mL SNA (B-1305) and MAL I (B-1315) to evaluate neuraminidase efficiency. To detect positive binding against the target of interest, cells were stained with anti-MUCl and anti-CD43 scFv antibodies at 5, 1.25, 0.3 and 0.08 pg/mL for 30 min on ice. Rabbit anti - MUC1 (HMFG2) (Abeam, ab245693) at 1 pg/mL and CD43-FITC (Miltenyi, 130-097-360) at 1:20 was used as a positive control. Cells were washed 2 times and stained with streptavidin Alexa Fluor 488 conjugated streptavidin (Invitrogen, S32354) at 1: 1000 to detect lectin binding, anti His Alexa Fluor 647 conjugated (R&D IC050R) at 1: 1000 to detect scFv binding, and goat ant-rabbit Alexa Fluor 647 (1: 1000) for 20 min on ice in the dark. After 2 washes, cells were analysed on Miltenyi Biotech-MACS Quant 16. Data analysis was performed using FlowJo Version 10. All lectins were supplied by Vector Biolabs.
2.7 Biolayer Interferometry (BLI)
Steady state kinetics were determined using an Octet Red96 system. Samples and buffers were dispensed into polypropylene 96well black flat-bottom plates (Greiner Bio- One, 655209) at a final volume of 200 pL per well, and all measurements were performed at 30°C with agitation at 1000 rpm. Prior to each assay, high streptavidin biosensor tips (SAX) (Sartorius, 18-5117) were prewetted in kinetics buffer (DPBS supplemented with 0.1% BSA and 0.02% Tween20) for at least 10 min followed by equilibration in kinetics buffer for 60s. The streptavidin biosensor tips were loaded with the biotinylated target glycopeptide in kinetics buffer for 300s, followed by an additional equilibration step of 100s. Association of scFvs in a range of different concentrations was performed for 300s. Finally, the dissociation was monitored with kinetics buffer for 300s. The association and dissociation responses were processed with the Octet Software (Version 12). Interferometry data was globally fitted to a 2: 1 model calculating the affinities and rate constants.
2.8 VL diversity sequencing
To follow phage enrichment during rounds of panning, the following LV diversity screening was performed. scFv VL-sequences that are related to the target peptide will appear with higher frequencies (enriched) confirming that phages with scFv are binding to the selected targets. These data is then be used for comparison with sequences from selected clones.
Sequencing was performed with Oxford Nanopore Technology (ONT). After each selection round, bacteria were scraped from the agar plate and an aliquot of the homogenous suspension was used for DNA purification using the GeneJet Miniprep Kit according to the manufacture's protocol. For the unselected libraries, homogenous suspension of scraped bacteria were used for DNA purification using Nucleobond Xtra EF Plasmid purification (MACHEREY-NAGEL GmbH & Co, 740422.50M) according to the manufacture's protocol. The set of primers that were used are found in the sequence listing (SEQ ID NOs. 50-53). Three pg of plasmid DNA were used as input material. Nanopore sequencing and data analysis was performed according to Karst et al 2021 with the following modifcations: a 0.8 x volume of AMPure XP beads was used for DNA clean-up after early and late PCR, all the DNA washes for the purification were performed with 80% ethanol. DNA was quantified using the Qubit dsHS DNA assay (Thermo Fisher Scientific). After late PCR, a 1% agarose gel was performed to verify the correct product size. Samples prepared for the R9 flow cell the SQK-LSK110 ligation sequencing kit protocol was used while samples prepared for the RIO flow cell the SQK-LSK114 ligation sequencing kit protocol was used. Samples run on a flow cell and sequencing was performed on a MinlON MklB device for 72 h.
Example 3: MUC1 as a proof of concept
MUC1 was used as proof of concept for the constructed libraries to identify binders against MUC1. Comparing the newly identified scFvs and the known mAbs in terms of sequence and antibody activity, the effectiveness and the functionality of both libraries was determined. The identified scFv were sequenced followed by VL chain analysis and characterized for their specificity on ELISA, cell binding assays and kinetic studies with BLI.
3.1 Phage selection
Tn template library, wherein each antibody in the library comprises G2D11 VH domain (SEQ ID NO. 1) was subjected to three rounds of selections by immobilizing target peptide 1 (see Table 1) on streptavidin-coated beads using target antigens. After each round of selection, phages were eluted, amplified and precipitated. Phage stocks after each round were titrated and analysed on polyclonal phage ELISA (Figure 5). Polyclonal phage ELISA showed specific binder enrichment for MUC1 target peptide in every round with zero non-specific binders against streptavidin. Interestingly, in round two and three binders against peptide 3 (see Table 1) have been enriched and therefore the library can be potentially used to identify binders for peptides with one GalNac. In addition, nanopore sequencing after each biopanning round was performed to verify sequence enrichment specific for MUC1. From previous studies, based on the crystallography of 5E5 mAb (SEQ ID NO. 3+4) and the other known mAbs, CDR3 sequence and specifically the motif YXY in CDR3 is responsible for MUC1 peptide backbone binding. Based on that observation, the twenty most frequent CDR3 sequences were ranked and sequence enrichment was demonstrated (Table 2). In addition, the MUC1 specific motif YXY in CDR3 sequences was identified.
*SEQ ID NOs 79-88 in sequence listing.
To specifically isolate binders for each target, DNA polyclonal phagemid pool after the third round was purified and the pool of different scFv genes were subcloned in the expression vector followed by individual scFv expression in a 96-well format. Expression of scFv in the supernatant was assessed with dot blot analysis (Figure 6). Sixty-one clones were picked and screened with monoclonal ELISA against target peptide 1 (see Table 1), and control peptides 4 and 8 (Figure 7). Peptide 4 was used to double confirm that with the Tn template library, mAbs against the backbone cannot be selected. The control peptide that was used in this study is IgA (see Table 1) that is produced in mucosal membranes and plays a significant role in their immunity. It has N- and Clunked glycosylation sites and is involved in a number of pathological conditions such as IgA deficiency and IgA nephropathy. Clones that showed no reactivity against the control peptides were selected for further characterization. In addition, sequences of the sixty one clones with sanger sequencing were obtained and alignment with VL sequences of 5E5 and 2D9 showed the differences in the CDR of the VL chains. Key binding features as presented in Table 2 were also present in VL-sequences from 5E5 and 2D9 further corborating their importance for interactions with the peptide backbone and determination of specificities.
3.2 Soluble scFv expression and evaluation
Based on monoclonal scFv ELISA specificity and VL sequence comparison, 10 clones were selected for purification. Of these ten clones, three clones (D3, A3 and D2) were fully evaluated and are represented by SEQ ID NOs. 9-11 in the sequences listing. Only their VL-domain is listed for the scFv. VH domain is same for the scFV of all clones - i.e. the VH of G2D11 (SEQ ID NO 1). The VL and VH are joined by the peptide linker (GGGS)5.
The selected clones were produced in larger scale for evaluation in ELISA, kinetic studies and cell binding assays. The scFvs were expressed and purified by His- tag affinity purification. Purity of the scFvs was checked on SDS-PAGE Coomassie analysis and western blot to confirm the presence of His-tag. For titration ELISA, soluble scFvs were screened for their binding at fixed concentration of MUC1 target peptide 1 and control peptide 8 (see table 1) Results are found in Figure 8A and 8B. In addition, scFvs were screened for cross-reactivity on other MUC1 peptides 2, 3, 4 and 5 (see Table 1). Results are found in Figure 9. Negative binding on unglycosylated MUC1 confirmed the library hypothesis that scFvs against the peptide backbone only cannot be selected. scFv D5 showed binding to monoTn-MUCl glycopeptide for concentration > 10 nM. Interestingly, clone D5 bind to peptide 3, while clones H3 and D3 also bind to peptide 3, but only at high concentrantion.
The clones that showed higher specificity for the target peptide 1 and zero crossreactivity to IgAl hinge region were chosen for further assessment. A3 and D2 clone demonstrated no binding to IgAl hinge region while D3 and H3 demonstrated binding at high antibody concentrations. The rest of the scFvs recognized also IgAl in addition to MUC1. Based on scFv specificity on titration ELISA A3, D2 and D3 scFvs were chosen for biological evaluation and kinetic studies.
For cell binding assays, two breast adenocarcinoma cell lines were employed, MCF7 and MDA-MB-231 WT and COSMC KO cells. COSMC KO means a knock-out of the COSMC gene which is a chaperon required to help catalyze the transfer of Galbl-3 to the penultimale sugat Tn (GalNAc), generating the Tn-antigen and subsequent elongation. In cancer this COSMC is a frequent phenomenon and the results of exposure of Tn and STn-antigens on tumor proteins. mAb HMFG2 (anti-MUCl) was used as a positive control to confirm MUC1 expression on cells (data not shown) and 5E5 was used as control mAb to confirm the Tn glycoform presence on MUC1. In this assay, when binding occurs, the the main peak in the flow cytometry diagram shifts to the right compared to the controls. All three scFv A3, D2 and D3 showed positive binding on MDA-MD-231 COSMC KO cells and negative binding on WT cells (Figure 8C). However, no positive binding was detected on MCF7 cells either for the selected MUC1 scFv clones or for 5E5 mAb control (Figure 10A). Neuraminidase treatment did not enhance binding of 5E5 control and selected MUC1 scFv clonesD3, D2 and A4 on MCF7 cells (data not shown).
To exclude non-specific binding, scFv binding to HEK293 was tested since they do not express endogenously MUC1. No binding was detected neither before nor after neuraminidase treatment (Figure 10B).
To further strengthen the concept of VL contribution in combotope specific recognition, the binding affinity (KD) of G2D11 and the newly identified MUC1 scFv clones were determined with BLI. To determine the on- and off-rates to calculate a dissociation constant, the biotinylated target peptide 1 was immobilized on streptavidin sensor tips. The only clones that showed higher binding affinity compared to G2D11 was clone D3 (Table 3).
Interestingly, the curves did not fit to a 1: 1 binding model but instead in a 2: 1 binding model (heterogeneous binding model). Therefore steady state kinetics cannot be calculated and in fact two KD values, KD and KD2, were obtained. One explanation for this is that VH-G2D11 binds bisTn in two orientations which makes it also more flexible in binding to Tn-structures in contrast to e.g. 5E5 that prefer a mono-Tn attached to a threonine (Thr) and less bis-Tn structures.
Based on the above, it is concluded that clones with peptide binding (combotopes) get an increase affinity (p-nM) compared to G2D11 (60nM). More antigen-antibody bonds interact to improve specificity and increase affninity.
3.3 Sequence analysis
Individual scFv sequence analysis revealed information about the CDR regions on the VL chains. scFvs D3, D5, H3 shared the MUC1 specific binding motif YSY in CDR3 like in 5E5 which is a requirement for peptide backbone binding interactions as it has been showed previously with crystallography (Figure 11). More specifically, Y98L and Y100L are contributing to the peptide binding. scFvs A3 and D2 shared the motif WNY while scFvs B5 and A2 had the motif SSY (Table 4). In all scFvs as well as the 5E5 mAb the Y100Lis conserved. CDR1 and CDR2 has some variations that can be linked to the target peptide sequence and size (e.g W50L) residue in cloe proximity to the peptide backbone.
*SEQ ID NOs 89-96 in sequence listing.
The obtained data correlate with current known information (x-ray and interacting residues, with specificities) from reference mAb 5E5. It demonstrates that the present library concept can enrich and select sequences with same features as obtained with immunized mice and hybridoma targeting the same antigen. This is a major step forward and with an animal free rapid system. The approach also provides a very large number of additional clones for evaluation with similar or different VL-sequence oprions that potentially could be better or different binders. The methods provides a relative (to hybridoma) controlled and systematic approach identifying large numbers of candidates for evaluation.
3.4 Binding profiles of the selected novel scFvs compared with known anti-Tn antibodies
Glycopeptides from Table 5 were printed on microarray chip, and the binding to these petides by scFv D3 and scFv A4 was compared with binding by scFv 5E5, scFv 2D9Chi, and scFv G2D11. 2D9Chi comprises the 2D9 VL domain and G2D11 VH domain. Results are presented in Figure 12 (for simplicity, the heat map shows amino acids 9-19 of the peptides in Table 5).
It was found that substitution of Thr with Ala as well as VTSA epitope abolished svFV 5E5 binding. Same epitope profile is demonstrated by scFv D3. It was shown that mono- Tn is not sufficient to get scFv G2D11 to bind, but scFv G2D11 binds to two adjacent Tn antigens. As previously disclosed herein, it is the VH domain of scFv G2D11 which is responsible for Tn binding. It was now further shown in this binding study, that if peptide binding is associated (such as in scFV D3 - i.e. having VH domain of G2D11 and a VL domain found in the library screening), then the VH domain of G2D11 can support monoTn binding.
ScFv 2D9Chi binds to two adjacent Tn antigens only in Ser-Thr sequence, contrary to scFV G2D11 which also binds two adjacent Tn antigens in the Thr-Ser sequence. ScFv A3 showed the same epitope recognition as 2D9Chi. Exchange of Pro residue with Ala, abolishes the binding of all scFvs against the glycopeptide.
Example 4: Concept evaluation - CD43 as first example
To assess the potentials of the Tn template library beyond MUC1, CD43 was chosen as the first target to identify binders following the same procedure.
Target peptide 9 (see table 1) was immobilized on streptavidin beads and three selection rounds were performed. To determine efficient phage selection, polyclonal phage ELISA and nanopore sequencing took place. Both polyclonal phage ELISA (Figure 13) and sequencing confirmed phage and sequence enrichments between the rounds. However, in the case of CD43 the sequences were grouped based on the combinations of CDR1, CDR2 and CDR3 as it is not evident which CDR is responsible for peptide backbone binding as in the case of MUC1 (Table 6).
*SEQ ID NOs 97-126 in sequence listing.
61 clones were picked, assessed for their binding on monoclonal scFv ELISA against target peptide 9 and control peptides 7 and 10 (see Table 1) (results in Figure 14) and VL sequences were obtained. Based on the monoclonal ELISA results ten clones that showed high specificity for the target peptide and zero or low cross reactivity were selected for further evaluation.
The ten clones (named ori, Hl, -Al, F4, C5, A7, D3, G3, D7, H2) were tested for their binding specificity on titration ELISA as a first step. These ten clones are represented by SEQ ID NOs. 12-21 in the sequences listing. Only their VL-domain is listed for the scFv. VH domain is same of the scFV of all clones - i.e. the VH of G2D11 (SEQ ID NO 1). The VL and VH are joined by the peptide linker (GGGS)s.
All scFvs showed high binding specificity for the CD43 target peptide (peptide no. 9), except C5 that showed the lowest binding specificity (Figure 15A). A7 and D3 clone showed high binding affinity for IgAl hinge region peptide, while Al and F4 showed some cross-binding only at high concentration (Figure 15B). The rest of the scFvs demonstrated no binding to IgAl peptide. scFvs Al, D7, Hl and H2 were chosen for further assessment in cell binding assays and kinetic studies with BLI. For biological evaluation, the leukemia Jurkat cells were chosen as they express high level of Tn antigen as a single nucleotide base deletion results in a frameshift and truncation of COSMC chaperone. Cells were treated with neuraminidase and stained in both cases. All of the four clones stained positively the Jurkat cells and the staining was enhanced upon neuraminidase treatment (Figure 15C). scFvs were tested additionally on HEK293 cells as they do not express naturally CD43 (Figure 16).
Finally, to confirm the VL contribution in the peptide binding, kinetic studies were performed with BLI. As in the case of MUC1 scFvs clones, the binding curves were fitted in a 2: 1 binding model (heterogeneous ligand) and two KD were obtained (Table 7).
Sequence analysis of the selected nine scFv gave information about the CRD1 and CDR3 region of the VL chains, as demonstrated by X-ray. X-ray experiments of ori-CD43 scFv (SEQ ID NO. 12) with CD43- GAS*T*GSP peptide (where asterisk denotes a GalNac residue) (SEQ ID NO. 56), revealed Y99L is required for specific peptide backbone binding interaction (Figure 17). Interestingly, all scFvs had the Y99L apart from D3, D7 and Hl. D3 scFv possess a L99L while D7 and Hl a W99L (Table 8). D3 with the L99L showed cross-binding to IgA hinge region peptide contrary to D7 and Hl that showed specific interaction with CD43 peptide and no cross-reactivity with TnlgA hinge or TnMUCl. However, scFv C5 that has a Y99L also cross reacted with the control peptide.
*SEQ ID NOs 127-134 in sequence listing.
With this CD43 example, a second target example was provided with a different peptide sequence but still mucin-like (amino acid features for O-glycosylation, high content of S, T, P, etc.), and enrichment of VL-domain sequences with a unique CDR fingerprint especially for CDR1 and CDR3 was demonstrated, and also confirmed by x-ray. The additional CDR1 interactions further increase specificity as demonstrated with ELISA (no cross-reactivity to TnlgA or TnMUCl). This demonstrates that the antibody library concept of the present invetion can target other Tn-peptide/proteins with different peptide sequence and obtain VL-domain sequences with corresponding signatures but different from above VL-domain sequences that targets Tn-MUCl.
Example 5: STn-template library
5.1 Alignment of G2D11 with 3F1
Antibody 3F1 is a known anti STn antibody (Prendergast et al 2017). Sequence comparison between G2D11 VH domain (SEQ ID NO. 1) and 3F1 VH domain (SEQ ID NO. 25) (see Figure 18) identified amino acid positions in CDR1 and CDR3 domains of G2D11 that are potentially responsible for STn glycan binding - specifically modifying the amino acid residues of G2D11 VH domain as follows: I28T, A30T, P101L, de/G102, T103A and F104L, seemed promising for changing the Tn glycan binding specificity of VH-G2D11 to STn glyan binding specificity.
5.2 G2D11 mutant with STn binding
G2D11 VH-domain mutants were prepared based on the above identified amino acid positions potentially relevant for STn specificity:
Ml (LAL): P101L, de/G102, T103A and F104L;
M2 (LAL-TFT): I28T, A30T, P101L, de/G102, T103A and F104L;
M3 (LAL-TFT-G): I28T, A30T, D56G, P101L, de/G102, T103A and F104L; and
M4 (TFT-G): I28T, A30T, and D56G. Microaray data for binding to glycopeptides by scFV G2D11, 3F1, and mutants (Ml-4). is illustrated in Figure 19. The mutants M1-M4 are mutants of scFV G2D11 comprising the above mentioned selected mutations in the VH domain. It was found that for the LAL mutant Ml, Tn binding was lost but no STn binding observed; while LAL+TFT mutant M2 generated STn binding and no Tn binding. These VH-domain mutations are sufficient to switch Tn to STn binding. Note that LAL is a requirement for the shift. No LAL mutation (TFT mutation only), Tn remains as a binder.
Hence, it was found that combination of mutation in VH chain, more precisely in CDR1 (IFA to TFT) and in CDR3 (PGTF to LAL), shifted the binding capacity from Tn to STn glycoform (Figure 19).
This, combined with the previous VH sequence alignment (Example 1.2) shows that the following amino acid residues are required VH residues to have binding towards bisSTn O-glycans are T28, T30, H32, A33, H35, Y50, S52, N55, D57, S99, L101, A102 and L103, with reference to SEQ ID NO 28.
5.3 STn-template library
Chain shuffling the mutated VH chain (SEQ ID NO. 28) (i.e. the G2D11 VH domain comprising the IFA to TFT mutation in CDR1 and PGTF to LAL in CDR3) with a pool of naive VL domains from naive mice a second phage display library was generated, termed the STn-template library.
Example 6: MUC1 as proof of concept in STn template library
To investigate the potentials of the STn template library, MUC1 was used a proof of concept. Two peptides, target peptide 6 and 7 (see Table 1), were immobilized on NHS beads and three selection round were performed as described previously. Nanopore sequencing after each round took place and the VL diversity between the two targets peptides was compared.
Table 9 and Table 10 provides the top ten most enriched combinations and their percentages in every round for bisSTn-MUCl and monoSTn-MUCl, respectively. In both selections, specific MUC1 binding motif Tyr-X-Tyr can be identified. The VL-domain sequences confirm that the STn-library can be used, and generates similar data/clones as obtained for TnMUCl.
*SEQ ID NOs 135-144 in sequence listing.
*SEQ ID NOs 145-154 in sequence listing. The enriched VL-domain sequences contains the same features as seen for Tn-MUCl and further consolidate the fingerprint related to MUC1 peptide target.
Based on monoclonal scFv ELISA specificity and VL sequence comparison, three clones were selected for purification. These three clones (named C4, D3, C7) are represented by SEQ ID NOs. 22-24 in the sequences listing. Only their VL-domain is listed for the scFv. VH domain is same for the scFV of all clones - i.e. the mutated VH chain of G2D11 (SEQ ID NO. 28) (i.e. the G2D11 VH domain comprising the IFA to TFT mutation in CDR1 and PGTF to LAL in CDR3, as disclosed in example 5). The VL and VH are joined by the peptide linker (GGGS)s.
Microarray analysis (Figure 20) showed that both bis-STn-MUCl and mono-STn-MUCl binders, but not Tn-binders were obtained.
Example 7: Kinetic affinities for MUC1 and CD43 specific scFvs
Table 11 summarizes the kinetic affinities for the MUC1 and CD43 specific scFvs indentified using the antibody libray according to the present invention.
Example 8: mono Tn/STn scFv binders
Enrichment of mono Tn MUC1 binders in the bisTn MUC1 biopanning as it was shown in the polyclonal phage ELISA (Figure 5, example 2) indicated that the library can also be panned with peptides with single GalNac attached to them. Indeed, MUC1 target peptides 2 and 3 (see table 1) were used to perform three rounds of selection. Polyclonal phage ELISA did not show phage enrichment, however sequencing of picked clones shared specific MUC1 sequences.
This is an important observation and may - without wishing to be bound by theory - be explained as follows: First, we know that monoTn-peptides are not binding to G2D11 (at least not without VL-contribution, see our ref Persson et al 2017). Mutations of VH (as stated above in example 1) also confirms this. G2D11 VH has high affinity (60nM) for bisTn contraty monoTn. This also leads to that panning with monoTn provides much less enrichment of clones as many of them are bisTn-binders. However, if quantity decreases, the quality increases and whatever binders that remain are most likely combotope binders with VL-peptide contribution to increase binding affinities. This is an advantage and demonstrates that the stringency can be influenced by manipulating the VH-domain binding strength either by removing one Tn or provide Tn-inhibitor or mutate the VH so it becomes less specific for bisTn, other panning conditions etc. Thus, we could see that selected scFv clones evaluated has high ratio of combotope binders and very little of bisTn-hapten binders. The nanopore sequence enrichment phage data also confirms this with monoSTn-MUCl. In contrast, enrichment with bisTn/STn-binders has much higher hapten binders and less combotope binders. All together, the antibody library concept of the present invention can be used to target monoTn/STn-peptide binders and not only bisTn/Stn-peptide binders, and this significantly increases utility as Tn/STn are situated as orfan, bis or in larger clusters.
Example 9: Humanizing mouse mAbs
A humanized scFv is generated by humanising the VL and VH immunoglobulin domains derived from the murine-originated antiCD43. Humanisation of VL and VH is performed in scFv format as follows:
Protocol:
1. Identify the complementarity-determining regions (CDRs): The CDRs are the parts of the antibody that interact with the antigen. The CDRs in the mouse monoclonal antibody are identified. Amino acid sequences of murine antiCD43 scFv VH and VL originated via phage display are numbered according to IGMT, and the IMGT-defined CDR residues are identified. Amino acids involved in binding but not included in the IMGT-defined CDR residues are also included as CDRs.
2. Design a humanized version of the antibody: Computational modeling tools are used to design a humanized version of the mouse monoclonal antibody. The goal is to maintain the antigen-binding specificity of the mouse antibody while replacing the mouse-derived CDRs with human-derived CDRs. From the amino acid sequence encompassing the 3 complementarity determining regions (CDRs) within the VL and VH domains, VH and VL sequences are generated in which the CDRs are masked. These sequences are used as input for the Basic Local Alignment Search Tool (BLAST) algorithm (Altschul et al., 1997) to identify similar frameworks from human V gene (heavy, kappa, lambda) germline databases. In addition, framework 4 amino acid sequence from the murine antiCD43 scFv VH and VL are used to identify similar human J gene segments.
Human V and J gene segments are chosen as template frameworks based on their identity to antiCD43 sequence, in-house analysis of individual and pairing frequency of V genes and previous experience of the use of particular templates for legacy humanisation.
The chosen human V gene frameworks are compared to the respective murine VH and VL sequences to identify potential sites that could undergo back-mutation to the corresponding mouse amino acid at that position. In-house collated evidences rules for the importance of certain framework positions in the likely maintenance of CDR conformation (and antigen binding affinity) are used to identify back-mutations considered most significant (primary mutations) and those of lower significance (secondary mutations). The extent of spatial clustering of the identified back-mutations is examined by analysing the crystallized molecular structure of mouse antiCD3. Initial humanised VH and VL sequences are generated by constructing a straight graft of the mouse CDRs into the chosen human germline templates. The apparent spatial clustering of back-mutation sites is used to reduce the potential number of variants of back- mutation containing humanised chains by introducing spatially-clustered mutations simultaneously.
Immunogenicity of the final humanised sequences is evaluated using online tools in AbYsis.
3. Clone the humanized antibody: The genes for the heavy and light chains of the humanized antibody are cloned into expression vectors. These vectors are used to produce the humanized antibody in a suitable expression system, such as mammalian cells. Amino acid sequences of humanised VH and VL chains are combined in the VL-VH orientation. scFv sequences comprised a (G4S)4 linker between VL and VH chains, and a C-terminal exa-His tag. scFv protein sequences are reverse translated and codon optimised.
All codon optimised DNA sequences are modified to include 5' and 3' adaptors suitable for HiFi cloning in the pET22b (+) and synthesised. DNA sequences are synthesised by TwistBioscience as double-stranded fragments (gBIocks). pET22b (+) backbone is linearized by PCR and the product is treated with Dnpl and cleaned with Monarch DNA&PCR cleanup. The gBIocks is inserted in pET22b (+) using the NEBuilder® HiFi DNA Assembly Cloning Kit. Ligation mixtures are transformed into DH5a competent E. coli and positive transformants selected on plates of LB agar supplemented with 100 ug/ml carbenicillin. Colonies for putative clones are cultured, plasmid DNA extracted, and DNA subjected to Sanger sequencing to identify correct clones.
Transient transfection of scFv-encoding construct into HEK 2936E suspension culture 250 pg of DNA for each plasmid construct is transfected using 293 Fectin transfection reagent into separate 250 ml cultures of HEK 293 6E cells (at a viable cell density of 1.85x10® cell/ml. The cultures are placed into a shaking 37°C incubator at 124rpm with 5% CO2. At 48 and 72 hours the cultures are supplemented with 6.2 ml tryptone (200 g/l) and 6.2 ml 3M fructose respectively.
From 48 hours post-transfection the viability (%) and viable cell density (cell/ml) of each culture are measured every 24 hours using a Vi-Cell cell counter and viability analyser (Beckman Coulter). Once the cultures have reached < 70% viability, the cultures are harvested via centrifugation at 4415xg for thirty minutes at 4°C and filtered via 0.22 pm Millipore filter. The filtered supernatants are stored at 4°C until required for protein purification.
4. Purify the humanized antibody: The humanized antibody is purified using Single- Step Affinity Protein Purification. The scFv proteins are purified from the resulting supernatants via AKTA Express system (AKTA). The supernatant is loaded at 5 ml/minute onto a 5 ml HisTrap Excel column pre-equilibrated with Buffer A (50 mM HEPES pH 7.5, 400mM NaCI, 20mM Imidazole). Once loaded, the column is washed in two column volumes of Buffer A at 5ml/minute back to baseline. The proteins are eluted in a step elution of 50% Buffer B (50 mM HEPES pH 7.5, 00 mM NaCI, IM Imidazole). The column is held in three column volumes of 50% Buffer B. During this step elution 0.5 ml fractions are collected in the purification of 88A, then 1 ml fractions are collected in all subsequent purifications. The elution step is continued until returned to baseline followed by a washout step at three column volumes of 100% Buffer B.
A single peak is expected at 280 nM on the resulting chromatogram indicating the elution of the protein of interest. The fractions corresponding to this peak are pooled and transferred to a 5000 MW cut-off centrifugal concentrator. The sample is buffer exchanged from Buffer B into 60ml PBS to separate the purified protein from the imidazole present in Buffer B. The samples are concentrated down to <1 ml. The concentration is determined via nanodrop and the purified protein is diluted in PBS to obtain a final concentration of 1 mg/ml. The final protein product is aliquoted and stored at -80°C for future use.
5pg aliquots of the purified samples in the batch are run under reducing conditions via SDS PAGE, displaying the purity and correct molecular weight of the purified proteins
5. Characterize the humanized antibody: The binding specificity and affinity of the humanized antibody to the target antigen is confirmed. Test for potential immunogenicity and other properties of the antibody such as stability and solubility. - Analytical Size Exclusion Chromatography (aSEC) for homogeneity of purified proteins
- Mass Spectrometry or Peptide Mass Fingerprinting to confirm protein identity.
- Biacore Analysis to confirm binding.
Example 10: humanization of mouse antibody clone D3
The objective of this example was to prepare a humanized clone of the mouse antibody scFv clone D3 (i.e. SEQ ID NO. 9 (VL domain) and SEQ ID NO. 1 (VH domain) joined by the peptide linker (GGGS)4).
Software used: igblastp 1.14.0, MAFFT v7.490, Sequence Manipulation Suite (DOI: 10.2144/00286ir01), and BioLuminate (Schrodinger).
10.1 Approach
1. A homology model of the parental mouse D3 Fv was built using Protein DataBank ID 2GKI, which provided the best template for VH/VL combined. Next, CDR loops were modeled, using a knowledge based approach resulting in 3 different models, each with other CDR templates.
Additionally, Alphafold v2 (local installation) was used to model D3 Fv from scratch.
2. Each model underwent a rigorous visual inspection and evaluation based on established protein quality metrics, such as those derived from Ramachandran plots. Among these, the model derived through homology from 2GKI showcased superior parameters. This model was subsequently employed to ascertain the significance of individual framework residues.
3. The parental mouse D3 sequence was used to interrogate the human germline sequence database with the BLASTP functionality integrated within Bioluminate. Additionally, maturated human antibody sequences were queried too. This resulted in 2 germline sequences for the VH: IGHV1-3 and IGHV1-69 and 1 germline sequence for the VL: IGKV4-1. Maturated sequences found were:
• Template PDB ID 7LKB: Crystal structure of PfCSP peptide 21 with vaccine- elicited human anti-malaria antibody m42.127 with 95.9% match human germline.
• Template PDB ID 7PS3: Crystal structure of antibody Beta-32 Fab with 94.9% match human germline.
• Additionally PDB ID 4LLU (pertuzumab) was considered, but resulted in more critical and potential critical mutations that this template was abandoned.
4. The templates were used to directly graft the CDRs based on IMGT definition with guidance of bioluminate. The resulting sequences were used to build homology models for future inspection. 5. The template sequences with grafted CDRs were aligned and residue by residue analyzed on potential (structural) issues.
6. Next, additional (back)mutations were introduced to ensure that CDR conformation will remain similar to the parental mouse Fv, humanness of framework based on IMGT definition was >80% (preferred >85%), potential sequence liabilities (e.g. post- translational modifications) were removed as best as possible.
7. For both VH and VL an additional framework was designed based on the consensus of the found templates, including (back)mutations as mentioned in previous point.
8. Sequences were renamed as D3VHx or D3VLx
9. The humanized VH and VL sequences are listed separate, to allow testing of different combinations - i.e. different combinations of D3VHx + D3VLx. Additionally, scFv sequences D3LxHx, with Gly-Ser linker were provided. Please note that signal peptide(s) are not provided.
10.2 Results - scFv sequences
The following combinations of D3VHx + D3VLx were found of particular interest: D3VL1 + D3VH1, D3VL1 + D3VH2, D3VL2+D3VH3, D3VL3 + D3VH4, and D3VL4 + D3VH5.
Hence, the following humanized scFv sequences were found of particular relevance: (The Gly-Ser linked proposed in (GGGGS)4, but could as well be (GGGGS)s.)
>D3LlHlscFv (SEQ ID NO. 168)
DIVMTQSPDSLAVSLGERATINCKSSQSLLYSSNQKNYLAWYQQKPGQAPKLLIYWASTRESGVP DRFSGSGSGTDFTLTISSLQAEDVAVYYCQQYYSYPLTFGAGTKLEMKGGGGSGGGGSGGGGS GGGGSEVQLVQSGAEVKKPGASVKVSCKASGYIFADHAIHWVRQAPGQRLEWIGYISPGNDDIK YNQKFQGRVTLTADKSASTAYMELSSLRSEDSAVYFCKRSLPGTFDYWGQGTTLTVSS
>D3LlH2scFv (SEQ ID NO. 169)
DIVMTQSPDSLAVSLGERATINCKSSQSLLYSSNQKNYLAWYQQKPGQAPKLLIYWASTRESGVP DRFSGSGSGTDFTLTISSLQAEDVAVYYCQQYYSYPLTFGAGTKLEMKGGGGSGGGGSGGGGS GGGGSEVQLVQSGAEVKKPGSSVKVSCKASGYIFADHAIHWVRRAPGQGLEWIGYISPGNDDIK YN E KFKG RATLTAD KSTSTAYM E LSS LRS E DTAVYFCKRS LPGTFDYWGQGTTLTVSS
>D3L2H3scFv (SEQ ID NO. 170)
DIVMTQSPDSLAVSLGEKATINCKSSQSLLYSSNQKNYLAWYQQKPGQPPKLLIYWASTRESGVP DRFSGSGSGTDFTLTISSLQAEDVAVYYCQQYYSYPLTFGGGTKVEIKGGGGSGGGGSGGGGSG GGGSEVQLVQSGAEVKKPGASVKVSCKASGYIFADHAIHWVRQAPGQRLEWIGYISPGNDDIKY SQKFQDKVTLTADKSASTAYMELSSLRSEDTAVYFCKRSLPGTFDYWGQGTTVTVSS >D3L3H4scFv (SEQ ID NO. 171)
DIQMTQSPSSVSASVGDRLTITCRSSQSLLYSSNQKNYLAWYQQKPGKAPKLLIYWASSLQSGV
PSRFSGSGSGTDFTLTISSLKPEDFATYYCQQYYSYPLTFGQGTKVEIKGGGGSGGGGSGGGGS GGGGSEVQLVQSGAEVKKPGASVKVSCKASGYIFADHAIHWVRQAPGQRLEWIGYISPGNDDIK YSQ E FQG RVTLTAD KSASTAYM E LSS LRS E DS AVYFCKRS LPGTFDYWGQGTTLTVSS
>D3L4H5scFv (SEQ ID NO. 172)
DIQMTQSPDSLAVSLGERATINCKSSQSLLYSSNQKNYLAWYQQKPGQPPKLLIYWASTRESGVP
DRFSGSGSGTDFTLTISSLQAEDVAVYYCQQYYSYPLTFGQGTKVEIKGGGGSGGGGSGGGGSG
GGGSEVQLVQSGAEVKKPGASVKVSCKASGYIFADHAIHWVRQAPGQRLEWIGYISPGNDDIKY SQEFQGRVTLTADKSASTAYMELSSLRSEDSAVYFCKRSLPGTFDYWGQGTTLTVSS
10.3 Results - Elisa titration
Elisa titration of the humanized scFvs on coated Tn-MUCl and other Tn-proteins was performed. The results are found in Figure 21. All humanized clones bound to the Tn- MUCl and to control proteins at high concentrations, but not when titered out. D3LlH2scFv showed similar binding characteristics as to the parental_D3_TnMUCl.
10.4 Results - Kd calue for D3L1H2
Kd meassurements was obtained by using Streptavidin coated Tip-sensors (SAX) from Startorius using manufacturer conditions and reagent kit, see Figure 22. The results are presented in table 12.
Example 11: Binding of mouse antibody scFv clone D3 to different Tn-peptide targets
Elisa titration of different target Tn-peptides detected with mouse D3 scFvs on streptavidin coated plates. The results are found in Figure 23. Biotinylated Tn-MUCl peptide and Tn-MUC13 peptides. TnMUCl (column 12, figure 23) was positive as well as one specific glycopeptide of TnMUC13 (column 6, figure 23, with similar secuence as MUC1); other peptides were negative.
PREFERRED EMBODIMENTS OF THE INVENTION PREFERRED EMBODIMENT 1. An antibody library for in-vitro identification of a specific antibody which binds a tumor cell, wherein each antibody in said library comprises
(i) a first antibody domain which binds a Tn- and/or STn-carbohydrate epitope of a glycoprotein of said tumor cell, and
(ii) a second antibody domain selected from a repertoire of second antibody domains, wherein the repertoire of second antibody domains comprises one or more second antibody domains which binds a peptide epitope of said glycoprotein of said tumor cell, wherein said specific antibody is specific for a combination of said carbohydrate epitope and said peptide epitope of said glycoprotein.
PREFERRED EMBODIMENT 2. The antibody library according to PREFERRED EMBODIMENT 1, wherein the carbohydrate epitope is covalently linked to the peptide epitope.
PREFERRED EMBODIMENT 3. The antibody library according to PREFERRED EMBODIMENT 1 or 2, wherein the first antibody domain is a VH-domain, and wherein the second antibody domain is a VL-domain.
PREFERRED EMBODIMENT 4. The antibody library according to any one of PREFERRED EMBODIMENTS 1-3, wherein the first antibody domain is a VH-domain, wherein the second antibody domain is a VL-domain, and wherein the antibodies in the library are scFv wherein the VH-domain is linked to the VL-domain via a peptide linker.
PREFERRED EMBODIMENT 5. The antibody library according to any one of PREFERRED EMBODIMENTS 1-4, wherein the carbohydrate epitope is selected from mono-Tn, bis-Tn, mono-STn, bis-STn, and a combination of mono-Tn and mono-STn.
PREFERRED EMBODIMENT 6. The antibody library according to any one of claims 3- 5, wherein the VH domain is a mono- or bis-Tn-binding VH domain, and wherein the VH- domain comprises an amino acid sequence having at least 90% sequence homology to SEQ ID NO. 1 and comprising amino acid residues H32, A33, H35, Y50, S52, N55, D57, and S99, with respect to SEQ ID NO. 1.
PREFERRED EMBODIMENT 6. The antibody library according to any one of PREFERRED EMBODIMENTS 1-5, wherein the first antibody domain is a mono- or bis- Tn-binding VH domain, and wherein said VH-domain comprises an amino acid sequence having at least 90% sequence homology to SEQ ID NO. 1 and comprising amino acid residues H32, A33, H35, Y50, and S99 and/or amino acid residues S52, N55, and D57, with respect to SEQ ID NO. 1.
PREFERRED EMBODIMENT 7. The antibody library according to any one of claims 3- 5, wherein the VH domain is a mono- or bis-STn-binding VH domain, and wherein the VH-domain comprises an amino acid sequence having at least 90% sequence homology to SEQ ID NO. 28 and comprising amino acid residues T28, T30, H32, A33, H35, Y50, S52, N55, D57, S99, L101, A102 and L103, with respect to SEQ ID NO. 28.
PREFERRED EMBODIMENT 7. The antibody library according to any one of PREFERRED EMBODIMENTS 1-5, wherein the first antibody domain is a mono- or bis- STn-binding VH domain, and wherein said VH-domain comprises an amino acid sequence having at least 90% sequence homology to SEQ ID NO. 28 and comprising amino acid residues (i), T28, T30, H32, A33, H35, Y50, S99, L101, A102 and L103, (ii), T28, T30, S52, N55, D57, L101, A102 and L103, or (iii) T28, T30, H32, A33, H35, Y50, S52, N55, D57, S99, L101, A102 and L103, with respect to SEQ ID NO. 28.
PREFERRED EMBODIMENT 8. The antibody library according to any one of PREFERRED EMBODIMENTS 1-7, wherein the first antibody domain does not contribute to or interfere with binding any peptide epitope.
PREFERRED EMBODIMENT 9. The antibody library according to any one of PREFERRED EMBODIMENTS 1-8, wherein the repertoire of second anytibody domains is generated from a naive immune repertoire of VL-domains, an immunized immune repertoire of VL-domains, or a synthetically produced repertoire of VL-domains; preferably a naive immune repertoire of VL-domains from an animal, such as a mouse or human.
PREFERRED EMBODIMENT 10. The antibody library according to any one of PREFERRED EMBODIMENTS 1-9, wherein the antibody library is a phage display library.
PREFERRED EMBODIMENT 11. A method for identifying an antibody for targeting a tumor cell, comprising the steps of i) preparing an antibody library according to any one of PREFERRED EMBODIMENTS 1- 10, and ii) screening said library to identify one or more tumor targeting antibodies.
PREFERRED EMBODIMENT 12. A method for identifying an antibody for targeting a glycoprotein, e.g. a glycoprotein on a tumor cell, comprising the steps of i) preparing an antibody library according to any one of PREFERRED EMBODIMENTS 1- 10, and ii) screening said library to identify one or more tumor targeting antibodies.
PREFERRED EMBODIMENT 13. The method according to PREFERRED EMBODIMENTS 11 or 12, comprising biopanning of the antibody library using a glycopeptide or glycoprotein of said tumor cell, prefarebly an O-glycosylated peptide or protein, such as Tn-mucin or other O-glycosylated protein having mucin-like motif, wherein said glycopeptide or glycoproteain is used in prurified form or expressed on a cell surface or tissue.
PREFERRED EMBODIMENT 14. The method according to any one of PREFERRED EMBODIMENTS 11-13, wherein the antibody library is a phage display library prepared by a method comprising the steps 1) mRNA isolation from a spleen, 2) cDNA synthesis from said mRNA, 3a) amplification from said cDNA using a specific set of primers to obtain a first nucleic acid sequence encoding the VH domain, 3b) application from said cDNA using a mix of primers to obtain multiple nucleic acid sequences encoding the repertoire of VL domains, 4) assembly of the first nucleic acid sequence endocing the VH-domain and a second nucleic acid sequence from the multiple nucleic acid sequences ecoding the VL-domain repertoire, to form a joint contruct, 5) insertion of the construct into a phagemid vector, 6) insertion of the phagemid vector comprising the construct into E. coli to produce a bacterial library, 7) using the bacterial library for infection of phages to produce the phage display library.
PREFERRED EMBODIMENT 15. A method for identifying a glycopeptide target, said target comprising a Tn and/or STn epitope and a peptide epitope, such as a glycopeptide target of a cancer cell, said method comprising the steps of i) preparing an antibody library according to any one of PREFERRED EMBODIMENTS 1- 10, ii) incubating the antibody library with a sample comprising the glycopeptide target, and iii) analyzing one or more antibody-peptide complexes obtained from step ii) to identify the amino acid sequence of the peptide epitope of said glycopeptide target.
PREFERRED EMBODIMENT 16. A specific tumor cell binding antibody, comprising
(i) a VH-domain which binds a Tn- and/or STn-carbohydrate epitope of a glycoprotein of the tumor cell, and
(ii) a VL-domain which binds a peptide epitope of said glycoprotein of said tumor cell, wherein the antibody is specific for a combination of said carbohydrate epitope and said peptide epitope of said glycoprotein, with the provisio that the antibody is not 5E5, 5F7 or 2D9.
PREFERRED EMBODIMENT 17. The antibody according to PREFERRED EMBODIMENT 16, wherein the VH domain comprises
(i) a first amino acid sequence having at least 90% sequence homology to SEQ ID NO. 1, wherein the first amino acid sequence comprises amino acid residues H32, A33, H35, Y50, S52, N55, D57, and S99, with respect to SEQ ID NO. 1; or
(ii) a second amino acid sequence having at least 90% sequence homology to SEQ ID NO. 28, and wherein the second amino acid sequence comprises amino acid residues T28, T30, H32, A33, H35, Y50, S52, N55, D57, S99, L101, A102 and L103, with respect to SEQ ID NO. 28.
PREFERRED EMBODIMENT 18. The antibody according to PREFERRED EMBODIMENT 16, wherein the VH domain comprises
(i) a first amino acid sequence having at least 90% sequence homology to SEQ ID NO. 1, wherein the first amino acid sequence comprises amino acid residues H32, A33, H35, Y50, and S99 and/or amino acid residues S52, N55, and D57, with respect to SEQ ID NO. 1; or
(ii) a second amino acid sequence having at least 90% sequence homology to SEQ ID NO. 28, and wherein the second amino acid sequence comprises amino acid residues Ci), T28, T30, H32, A33, H35, Y50, S99, L101, A102 and L103, (ii), T28, T30, S52, N55, D57, L101, A102 and L103, or (iii) T28, T30, H32, A33, H35, Y50, S52, N55, D57, S99, L101, A102 and L103, with respect to SEQ ID NO. 28.
PREFERRED EMBODIMENT 19. The antibody according to any one of PREFERRED EMBODIMENTS 16-18, wherein the VL domain comprises an amino acid sequence selected from SEQ ID NO. 9-24.
PREFERRED EMBODIMENT 20. The antibody according to any one of PREFERRED EMBODIMENTS 16-19, comprising a first and a second antigen-binding fragement comprising a first VH and VL domain and a second VH and VL domain, respectively, wherein the first VH domain binds a Tn-carbohydrate epitope and the second VH domain binds a STn carbohydrate epitope.
PREFERRED EMBODIMENT 21. The antibody according to any one of PREFERRED EMBODIMENT 20, wherein the first VH domain comprises a first amino acid sequence having at least 90% sequence homology to SEQ ID NO. 1, wherein the first amino acid sequence comprises amino acid residues H32, A33, H35, Y50, and S99 and/or amino acid residues S52, N55, and D57, with respect to SEQ ID NO. 1; and wherein the second VH domain comprises a second amino acid sequence having at least 90% sequence homology to SEQ ID NO. 28, and wherein the second amino acid sequence comprises amino acid residues (i), T28, T30, H32, A33, H35, Y50, S99, L101, A102 and L103, (ii), T28, T30, S52, N55, D57, L101, A102 and L103, or (iii) T28, T30, H32, A33, H35, Y50, S52, N55, D57, S99, L101, A102 and L103, with respect to SEQ ID NO. 28.
PREFERRED EMBODIMENT 22. The antibody according to any one of PREFERRED EMBODIMENTS 16-21 for use in treatment and/or prevention of cancer.
PREFERRED EMBODIMENT 23. A method of treating a cancer in a subject comprising administering a formulation comprising at least one antibody according to any one of PREFERRED EMBODIMENTS 16-21 to a patient in need thereof.
PREFERRED EMBODIMENT 24. The antibody according to any one of PREFERRED EMBODIMENTS 16-21 for use in diagnosing a cancerous condition in a subject.
PREFERRED EMBODIMENT 25. A method of diagnosing a cancerous condition in a subject comprising administering a formulation comprising at least one antibody according to any one of PREFERRED EMBODIMENTS 16-21 to the subject, and detecting the presence of an antigen-antibody complex comprising said at least one antibody according to any one of PREFERRED EMBODIMENTS 16-21 and a Tn- and/or STn- antigen.
REFERENCES
Blixt et al 2010. A High-throughput O-Glycopeptide Discovery Platform for
Seromic Profiling. J Proteome Res. 2010 October 1; 9(10): 5250-5261. doi: 10.1021/prl005229.
Blixt et al 2012. Analysis of Tn antigenicity with a panel of new IgM and IgGl monoclonal antibodies raised against leukemic cells. Glycobiology. 2012 Apr;22(4): 529- 42. doi : 10.1093/glycob/cwrl78. Epub 2011 Dec 5.
Clavero-Alvarez et al 2018. Humanization of Antibodies using a Statistical Inference Approach. Sci Rep. 2018 Oct 4;8(1): 14820. doi: 10.1038/s41598-018-32986-y.
Karst et al. 2021. High-accuracy long-read amplicon sequences using unique molecular identifiers with Nanopore or PacBio sequencing. Nat Methods 18, 165-169 (2021). Kjeldsen et al 1988. Preparation and characterization of monoclonal antibodies directed to the tumor-associated O-linked sialosyl-2 — 6 alpha-N-acetylgalactosaminyl (sialosyl-Tn) epitope. Cancer Res. 1988 Apr 15;48(8):2214-20.
Kudelka et al 2015. Simple Sugars to Complex Disease— Mucin-Type O-Glycans in Cancer. Adv Cancer Res. 2015; 126: 53-135. Doi: 10.1016/bs.acr.2014.11.002.
Li et al 2009. Resolving conflicting data on expression of the Tn antigen and implications for clinical trials with cancer vaccines. Mol Cancer Ther. 2009 Apr;8(4):971- 9. doi: 10.1158/1535-7163. MCT-08-0934. PMID: 19372570; PMCID: PMC2752371.
Macias-Leon et al 2020. Structural characterization of an unprecedented lectin-like antitumoral anti-MUCl antibody. Chem Commun (Camb). 2020 Dec 8;56(96): 15137- 15140. doi: 10.1039/d0cc06349e.
Mazal et al 2013. Monoclonal antibodies toward different Tn-amino acid backbones display distinct recognition patterns on human cancer cells. Implications for effective immuno-targeting of cancer. Cancer Immunol Immunother. 2013 Jun;62(6): 1107-22. doi: 10.1007/S00262-013-1425-7. Epub 2013 Apr 21. PMID: 23604173.
Oppezzo Pet al 2000. Production and functional characterization of two mouse/human chimeric antibodies with specificity for the tumor-associated Tn-antigen. Hybridoma. 2000 Jun; 19(3):229-39. doi: 10.1089/02724570050109620. PMID: 10952411.
Persson et al. 2016. A combinatory antibody-antigen microarray assay for high- content screening of single-chain fragment variable clones from recombinant libraries. PLoS One 11, (2016).
Person et al. 2017. Epitope mapping of a new anti-Tn antibody detecting gastric cancer cells. Glycobiology, 2017, vol. 27, no. 7, 635-645. doi: 10.1093/glycob/cwx033.
Prendergast et al 2017. Novel anti-Sialyl-Tn monoclonal antibodies and antibody-drug conjugates demonstrate tumor specificity and anti-tumor activity. MAbs. 2017 May/Jun;9(4):615-627. doi: 10.1080/19420862.2017.1290752. Epub 2017 Feb 22.
Serensen et al 2006. Chemoenzymatically synthesized multimeric Tn/STn MUC1 glycopeptides elicit cancer-specific anti-MUCl antibody responses and override tolerance. Glycobiology. 2006 Feb; 16(2):96-107. doi: 10.1093/glycob/cwj044. Epub 2005 Oct 5.
Tarp et al 2007. Identification of a novel cancer-specific immunodominant glycopeptide epitope in the MUC1 tandem repeat. Glycobiology vol. 17 no. 2 pp. 197-209, 2007. doi: 10.1093/glycob/cwl061.
Yuasa et al 2012. Construction and expression of anti-Tn-antigen-specific single-chain antibody genes from hybridoma producing MLS128 monoclonal antibody. J Biochem. 2012 Apr; 151(4):371-81. doi: 10.1093/jb/mvs007. Epub 2012 Feb 8. PMID: 22318767.

Claims

1. An antibody library for in-vitro identification of a specific antibody which binds a tumor cell, wherein each antibody in said library comprises
(i) a VH-domain which binds a Tn- and/or STn-carbohydrate epitope of a glycoprotein of said tumor cell, and
(ii) a VL-domain selected from a repertoire of VL-domains, wherein the repertoire of VL-domains comprises one or more VL-domains which binds a peptide epitope of said glycoprotein of said tumor cell, wherein said specific antibody is specific for a combination of said carbohydrate epitope and said peptide epitope of said glycoprotein.
2. The antibody library according to claim 1, wherein the carbohydrate epitope is covalently linked to the peptide epitope.
3. The antibody library according to claim 1 or 2, wherein the antibodies in the library are scFv wherein the VH-domain is linked to the VL-domain via a peptide linker.
4. The antibody library according to any one of claims 1-3, wherein the carbohydrate epitope is selected from mono-Tn, bis-Tn, mono-STn, bis-STn, and a combination of mono-Tn and mono-STn.
5. The antibody library according to any one of claims 1-4, wherein the VH domain is a mono- or bis-Tn-binding VH domain, and wherein the VH-domain comprises an amino acid sequence having at least 90% sequence homology to SEQ ID NO. 1 and comprising amino acid residues H32, A33, H35, Y50, and S99 and/or amino acid residues S52, N55, and D57, with respect to SEQ ID NO. 1.
6. The antibody library according to any one of claims 1-4, wherein the VH domain is a mono- or bis-Tn-binding VH domain, and wherein the VH-domain comprises an amino acid sequence having at least 90% sequence homology to SEQ ID NO. 1 and comprising amino acid residues H32, A33, H35, Y50, S52, N55, D57, and S99, with respect to SEQ ID NO. 1.
7. The antibody library according to any one of claims 1-4, wherein the VH domain is a mono- or bis-STn-binding VH domain, and wherein the VH-domain comprises an amino acid sequence having at least 90% sequence homology to SEQ ID NO. 28 and comprising amino acid residues (i), T28, T30, H32, A33, H35, Y50, S99, L101, A102 and L103, (ii), T28, T30, S52, N55, D57, L101, A102 and L103, or (iii) T28, T30, H32, A33, H35, Y50, S52, N55, D57, S99, L101, A102 and L103, with respect to SEQ ID NO. 28.
8. The antibody library according to any one of claims 1-4, wherein the VH domain is a mono- or bis-STn-binding VH domain, and wherein the VH-domain comprises an amino acid sequence having at least 90% sequence homology to SEQ ID NO. 28 and comprising amino acid residues T28, T30, H32, A33, H35, Y50, S52, N55, D57, S99, L101, A102 and L103, with respect to SEQ ID NO. 28.
9. The antibody library according to any one of claims 1-8, wherein the VH domain domain does not contribute to or interfere with binding any peptide epitope.
10. The antibody library according to any one of claims 1-9, wherein the repertoire of VL-domains is generated from a naive immune repertoire of VL-domains, an immunized immune repertoire of VL-domains, or a synthetically produced repertoire of VL-domains; preferably a naive immune repertoire of VL-domains from an animal, such as a mouse or human.
11. The antibody library according to any one of claims 1-10, wherein the antibody library is a phage display library.
12. A nucleic acid library encoding the antibody library of any one of claims 1-11.
13. A method for identifying an antibody for targeting a glycoprotein, comprising the steps of i) preparing an antibody library according to any one of claims 1-12, and ii) screening said library to identify one or more tumor targeting antibodies.
14. The method according to claim 13, comprising biopanning of the antibody library using a glycopeptide or glycoprotein of said tumor cell, prefarebly an O-glycosylated peptide or protein, such as Tn-mucin or other O-glycosylated protein having mucin-like motif, wherein said glycopeptide or glycoproteain is used in prurified form or expressed on a cell surface or tissue.
15. The method according to claim 13 or 14, wherein the antibody library is a phage display library prepared by a method comprising the steps 1) mRNA isolation from a spleen, 2) cDNA synthesis from said mRNA, 3a) amplification from said cDNA using a specific set of primers to obtain a first nucleic acid sequence encoding the VH domain, 3b) application from said cDNA using a mix of primers to obtain multiple nucleic acid sequences encoding the repertoire of VL domains, 4) assembly of the first nucleic acid sequence endocing the VH-domain and a second nucleic acid sequence from the multiple nucleic acid sequences ecoding the VL-domain repertoire, to form a joint contruct, 5) insertion of the construct into a phagemid vector, 6) insertion of the phagemid vector comprising the construct into E. coli to produce a bacterial library, 7) using the bacterial library for infection of phages to produce the phage display library.
16. A method for identifying a glycopeptide target, said target comprising a Tn and/or STn epitope and a peptide epitope, such as a glycopeptide target of a cancer cell, said method comprising the steps of i) preparing an antibody library according to any one of claims 1-11, ii) incubating the antibody library with a sample comprising the glycopeptide target, and iii) analyzing one or more antibody-peptide complexes obtained from step ii) to identify the amino acid sequence of the peptide epitope of said glycopeptide target.
17. A specific tumor cell binding antibody, comprising
(i) a VH-domain which binds a Tn- and/or STn-carbohydrate epitope of a glycoprotein of the tumor cell, and
(ii) a VL-domain which binds a peptide epitope of said glycoprotein of said tumor cell, wherein the antibody is specific for a combination of said carbohydrate epitope and said peptide epitope of said glycoprotein, with the provisio that the antibody is not 5E5, 5F7 or 2D9.
18. The antibody according to claim 17, wherein the VH domain comprises
(i) a first amino acid sequence having at least 90% sequence homology to SEQ ID NO. 1, wherein the first amino acid sequence comprises amino acid residues H32, A33, H35, Y50, and S99 and/or amino acid residues S52, N55, and D57, with respect to SEQ ID NO. 1; or
(ii) a second amino acid sequence having at least 90% sequence homology to SEQ ID NO. 28, and wherein the second amino acid sequence comprises amino acid residues (i), T28, T30, H32, A33, H35, Y50, S99, L101, A102 and L103, (ii), T28, T30, S52, N55, D57, L101, A102 and L103, or (iii) T28, T30, H32, A33, H35, Y50, S52, N55, D57, S99, L101, A102 and L103, with respect to SEQ ID NO. 28.
19. The antibody according to claim 17, wherein the VH domain comprises
(i) a first amino acid sequence having at least 90% sequence homology to SEQ ID NO. 1, wherein the first amino acid sequence comprises amino acid residues H32, A33, H35, Y50, S52, N55, D57, and S99, with respect to SEQ ID NO. 1; or (ii) a second amino acid sequence having at least 90% sequence homology to SEQ ID NO. 28, and wherein the second amino acid sequence comprises amino acid residues T28, T30, H32, A33, H35, Y50, S52, N55, D57, S99, L101, A102 and L103, with respect to SEQ ID NO. 28.
20. The antibody according to claim 17-19, wherein the VL domain comprises an amino acid sequence selected from SEQ ID NO. 9-24.
21. The antibody according to any one of claims 17-20, comprising a first and a second antigen-binding fragement comprising a first VH and VL domain and a second VH and VL domain, respectively, wherein the first VH domain binds a Tn-carbohydrate epitope and the second VH domain binds a STn carbohydrate epitope.
22. The antibody according to claim 21, wherein the first VH domain comprises a first amino acid sequence having at least 90% sequence homology to SEQ ID NO. 1, wherein the first amino acid sequence comprises amino acid residues H32, A33, H35, Y50, and S99 and/or amino acid residues S52, N55, and D57, with respect to SEQ ID NO. 1; and wherein the second VH domain comprises a second amino acid sequence having at least 90% sequence homology to SEQ ID NO. 28, and wherein the second amino acid sequence comprises amino acid residues (i), T28, T30, H32, A33, H35, Y50, S99, L101, A102 and L103, (ii), T28, T30, S52, N55, D57, L101, A102 and L103, or (iii) T28, T30, H32, A33, H35, Y50, S52, N55, D57, S99, L101, A102 and L103, with respect to SEQ ID NO. 28.
23. The antibody according to any one of claims 17-22 for use in treatment and/or prevention of cancer.
24. A method of treating a cancer in a subject comprising administering a formulation comprising at least one antibody according to any one of claims 17-22 to a patient in need thereof.
25. The antibody according to any one of claims 17-22 for use in diagnosing a cancerous condition in a subject.
26. A method of diagnosing a cancerous condition in a subject comprising administering a formulation comprising at least one antibody according to any one of claims 17-22 to the subject, and detecting the presence of an antigen-antibody complex comprising said at least one antibody according to any one of claims 17-22 and a Tn- and/or STn- antigen.
EP24707819.9A 2023-03-17 2024-03-01 Combotope antibody libraries Pending EP4680631A1 (en)

Applications Claiming Priority (3)

Application Number Priority Date Filing Date Title
PCT/EP2023/056922 WO2024193794A1 (en) 2023-03-17 2023-03-17 Combotope antibody libraries
US18/526,205 US20240309110A1 (en) 2023-03-17 2023-12-01 Combotope Antibody Libraries
PCT/EP2024/055483 WO2024193989A1 (en) 2023-03-17 2024-03-01 Combotope antibody libraries

Publications (1)

Publication Number Publication Date
EP4680631A1 true EP4680631A1 (en) 2026-01-21

Family

ID=90059313

Family Applications (1)

Application Number Title Priority Date Filing Date
EP24707819.9A Pending EP4680631A1 (en) 2023-03-17 2024-03-01 Combotope antibody libraries

Country Status (5)

Country Link
EP (1) EP4680631A1 (en)
JP (1) JP2026509914A (en)
CN (1) CN120882741A (en)
AU (1) AU2024238912A1 (en)
WO (1) WO2024193989A1 (en)

Family Cites Families (5)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
AU2007304590A1 (en) 2006-10-04 2008-04-10 Cancer Research Technology Limited Generation of a cancer-specific immune response toward MUC1 and cancer specific MUC1 antibodies
FI3218005T3 (en) 2014-11-12 2023-03-31 Seagen Inc GLYCAN INTERACTING COMPOUNDS AND METHODS OF USE
US11161911B2 (en) 2017-10-23 2021-11-02 Go Therapeutics, Inc. Anti-glyco-MUC1 antibodies and their uses
US20210060070A1 (en) 2019-09-04 2021-03-04 Tmunity Therapeutics Inc. Adoptive cell therapy and methods of dosing thereof
CN118354789A (en) 2021-09-03 2024-07-16 Go医疗股份有限公司 Anti-glycemic-cMET antibodies and their uses

Also Published As

Publication number Publication date
AU2024238912A1 (en) 2025-09-25
CN120882741A (en) 2025-10-31
JP2026509914A (en) 2026-03-25
WO2024193989A1 (en) 2024-09-26

Similar Documents

Publication Publication Date Title
TWI402078B (en) Antibodies against csf-1r
AU2021260639A1 (en) Antibody against Nectin-4 and application thereof
KR102257462B1 (en) Anti-PCSK9 monoclonal antibody
Ho et al. A novel high‐affinity human monoclonal antibody to mesothelin
KR20190134614A (en) B7-H3 antibody, antigen-binding fragment thereof and medical use thereof
CN108997499B (en) An anti-human PD-L1 antibody and its application
CN110590952B (en) High-affinity nano antibody for anti-CA 125 carbohydrate antigen and application thereof
CN113227148B (en) anti-GPC 3 antibody, antigen-binding fragment thereof, and medical use thereof
CA2938933A1 (en) Anti-laminin4 antibodies specific for lg4-5
WO2022247804A1 (en) Anti-gprc5d antibody, preparation method therefor, and use thereof
CN115505043A (en) Antibodies specifically binding glycosylated CEACAM5
CN101701039B (en) Variable regions of light chains and heavy chains of FMU-EPCAM-2A9 monoclonal antibodies
CN112480250B (en) Anti-human osteopontin antibody and application thereof
CN107108734B (en) Monoclonal anti-GPC-1 antibodies and uses thereof
EP3439692A2 (en) Plectin-1 binding antibodies and uses thereof
Cho et al. Generation, characterization and preclinical studies of a human anti-L1CAM monoclonal antibody that cross-reacts with rodent L1CAM
CN108290942B (en) Antibodies cross-linked with human and mouse semaphorin 3A and uses thereof
WO2024012434A1 (en) Antibody, antigen-binding fragment thereof, and pharmaceutical use thereof
US20240309110A1 (en) Combotope Antibody Libraries
EP4680631A1 (en) Combotope antibody libraries
TWI898089B (en) Antibody for enrichment of cells
GB2619976A (en) Humanised antibodies or functional fragments thereof against tumour antigens
US20130109586A1 (en) Generation of antibodies to an epitope of interest that contains a phosphomimetic amino acid
JP7498747B2 (en) Anti-GM2AP antibody and its applications
RU2761876C1 (en) Humanised 5d3hu antibody binding to the prame tumour antigen, dna fragments encoding said antibody, and antigen-binding fragment of the antibody

Legal Events

Date Code Title Description
STAA Information on the status of an ep patent application or granted ep patent

Free format text: STATUS: UNKNOWN

STAA Information on the status of an ep patent application or granted ep patent

Free format text: STATUS: THE INTERNATIONAL PUBLICATION HAS BEEN MADE

PUAI Public reference made under article 153(3) epc to a published international application that has entered the european phase

Free format text: ORIGINAL CODE: 0009012

STAA Information on the status of an ep patent application or granted ep patent

Free format text: STATUS: REQUEST FOR EXAMINATION WAS MADE

17P Request for examination filed

Effective date: 20251013

AK Designated contracting states

Kind code of ref document: A1

Designated state(s): AL AT BE BG CH CY CZ DE DK EE ES FI FR GB GR HR HU IE IS IT LI LT LU LV MC ME MK MT NL NO PL PT RO RS SE SI SK SM TR