WO2024258911A1 - Sequence specific dna binding proteins - Google Patents

Sequence specific dna binding proteins Download PDF

Info

Publication number
WO2024258911A1
WO2024258911A1 PCT/US2024/033519 US2024033519W WO2024258911A1 WO 2024258911 A1 WO2024258911 A1 WO 2024258911A1 US 2024033519 W US2024033519 W US 2024033519W WO 2024258911 A1 WO2024258911 A1 WO 2024258911A1
Authority
WO
WIPO (PCT)
Prior art keywords
seq
polypeptide
engineered
tag
mutation
Prior art date
Legal status (The legal status is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the status listed.)
Ceased
Application number
PCT/US2024/033519
Other languages
French (fr)
Inventor
Ruijie ZHANG
Derek Hunter VALLEJO
Katherine NAKAMA
Lorita BOGHOSPOR
Lauren GUTGESELL
Current Assignee (The listed assignees may be inaccurate. Google has not performed a legal analysis and makes no representation or warranty as to the accuracy of the list.)
10X Genomics Inc
Original Assignee
10X Genomics Inc
Priority date (The priority date is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the date listed.)
Filing date
Publication date
Application filed by 10X Genomics Inc filed Critical 10X Genomics Inc
Priority to EP24739890.2A priority Critical patent/EP4728062A1/en
Publication of WO2024258911A1 publication Critical patent/WO2024258911A1/en
Anticipated expiration legal-status Critical
Ceased legal-status Critical Current

Links

Classifications

    • CCHEMISTRY; METALLURGY
    • C12BIOCHEMISTRY; BEER; SPIRITS; WINE; VINEGAR; MICROBIOLOGY; ENZYMOLOGY; MUTATION OR GENETIC ENGINEERING
    • C12NMICROORGANISMS OR ENZYMES; COMPOSITIONS THEREOF; PROPAGATING, PRESERVING, OR MAINTAINING MICROORGANISMS; MUTATION OR GENETIC ENGINEERING; CULTURE MEDIA
    • C12N9/00Enzymes; Proenzymes; Compositions thereof; Processes for preparing, activating, inhibiting, separating or purifying enzymes
    • C12N9/10Transferases (2.)
    • C12N9/12Transferases (2.) transferring phosphorus containing groups, e.g. kinases (2.7)
    • C12N9/1241Nucleotidyltransferases (2.7.7)
    • C12N9/1276RNA-directed DNA polymerase (2.7.7.49), i.e. reverse transcriptase or telomerase
    • CCHEMISTRY; METALLURGY
    • C12BIOCHEMISTRY; BEER; SPIRITS; WINE; VINEGAR; MICROBIOLOGY; ENZYMOLOGY; MUTATION OR GENETIC ENGINEERING
    • C12YENZYMES
    • C12Y207/00Transferases transferring phosphorus-containing groups (2.7)
    • C12Y207/07Nucleotidyltransferases (2.7.7)
    • C12Y207/07049RNA-directed DNA polymerase (2.7.7.49), i.e. telomerase or reverse-transcriptase
    • CCHEMISTRY; METALLURGY
    • C07ORGANIC CHEMISTRY
    • C07KPEPTIDES
    • C07K2319/00Fusion polypeptide

Definitions

  • the specific position of a cell within a tissue can also affect the cell’s morphology, differentiation, fate, viability, proliferation, behavior, signaling, and cross-talk with other cells in the tissue.
  • This spatial heterogeneity has been previously studied using techniques that only provide data for a small handful of analytes in the context of an intact tissue or a portion of a tissue. Specifically, these techniques can provide substantial analyte data for dissociated tissues (i.e., single cells). However, they fail to provide information regarding the position of a single cell in a biological sample (e.g., tissue sample) caused in part by inefficient transcript capture.
  • RT polypeptide comprising: (a) an RT polypeptide sequence; (b) a DNA binding domain, where the DNA binding domain is from a molecule capable of binding a minor groove of a nucleic acid; and (c) a linker connecting the RT polypeptide sequence and the DNA binding domain.
  • the DNA binding domain is located at the N-terminus of the RT polypeptide sequence. In some embodiments, the DNA binding domain is located at the C-terminus of the RT polypeptide sequence.
  • the linker is a glycine-serine (GS) linker selected from the group consisting of (GS)n, (GSGGS)n, (SGGSG)n, (GGGS)n, (GGSG)n, (GGSGG)n, (GSGSG)n, (GSGGG)n, GGGSG)n, and (GSSSG)n, where n represents an integer of at least 1.
  • the linker is GGGS.
  • the linker is SGGSG.
  • the DNA binding domain specifically recognizes adenine-thymine-rich region on a nucleic acid molecule.
  • the DNA binding domain specifically recognizes oligo(dA) or oligo(dT) tracts on a nucleic acid molecule.
  • the DNA binding domain comprises at least one AT-rich interaction domain.
  • the DNA binding domain comprises at least 2, at least 3, at least 4, at least 5, at least 6, at least 7, at least 8, at least 9, or at least 10 AT-rich interaction domains.
  • the AT-rich interaction domain comprises a core sequence, wherein the core sequence is a two base core sequence, a three base core sequence, a four base core sequence, or a five base core sequence.
  • At least one of the bases of the core sequence comprises an arginine; a glycine and an arginine; a proline and an arginine; a lysine and an arginine; or any combination thereof.
  • the AT-rich interaction domain comprises a GRKPG (Gly-Arg-Lys-Pro-Gly) repeat, a RKRGRPKK repeat, a KKRGRPKK repeat, a RKRGR repeat, a GR*R/PPK repeat, a GR*RPK repeat, a GR*PPK repeat, a KRPR* repeat, or a K/RKRGRPKK repeat.
  • the AT-rich interaction domain comprises a core sequence comprising an amino acid selected from the group consisting of SEQ ID NO: 11- 24.
  • the DNA binding domain is a DNA binding domain of any one of Saccharomyces cerevisiae datin (DAT1), high mobility group AT hook 1 (HMGA1), lysine-specific methyltransferase 2a ( KMT2A), Myocyte Enhancer Factor 2C (MEF2C), Heterogeneous Nuclear Ribonucleoprotein D (HNRNPD), Structural Maintenance of Chromosomes 1A (SMC1), Structural Maintenance Of Chromosomes 2 (SMC2), Caenorhabditis elegans tbp-1, Drosophila melanogaster D1 protein, Salmonella typhimurium Hin recombinase, S.
  • DAT1 Saccharomyces cerevisiae datin
  • HMGA1 high mobility group AT hook 1
  • KMT2A lysine-specific methyltransferase 2a
  • MEF2C Myocyte Enhancer Factor 2C
  • HNRNPD Heterogeneous Nuclear
  • the DNA binding domain is from a S. cerevisiae DAT1.
  • the amino acid sequence of the DNA binding domain comprises a DNA binding domain consensus motif set forth in SEQ ID NO: 13, 14, 16, or 22.
  • the DNA binding domain comprises: (a) a full-length DAT1 sequence or SEQ ID NO: 2; (b) an N-terminal truncated variant of DAT1 (D90) comprising the first 90 amino acids of the full length DAT1, or SEQ ID NO: 3; (c) a truncated variant of DAT1 (D60) comprising the first 60 amino acids of the full length DAT1 or SEQ ID NO: 5; (d) a truncated variant of DAT1 (D48) comprising the first 48 amino acids of the full length DAT1 or SEQ ID NO: 6; (e) a truncated variant of DAT1(D36) comprising the first 36 amino acids of the full length DAT1 or SEQ ID NO: 8; (f) a truncated variant of DAT1 (D35) comprising the first 35 amino acids of full length DAT1 or SEQ ID NO: 9; (g) an amino acid sequence having at least about 90%,
  • the DNA binding domain comprises the amino acid sequence of SEQ ID NO: 11 (GRKPG). In some embodiments, the DNA binding domain optionally comprise at least 2 domains or at least 3 domains comprising SEQ ID NO: 11. [00015] In some embodiments, the DNA binding domain comprises a mutation in any of one of SEQ ID NO: 2, 3, 5, 6, 8, 9, or 11. In some embodiments, the mutation is selected from a substitution, an insertion, a deletion, or any combination thereof. [00016] In some embodiments of the engineered RT polypeptide described herein, the DNA binding domain comprises SEQ ID NO: 2. In some embodiments of the engineered RT polypeptide described herein, the DNA binding domain comprises the amino acid sequence of SEQ ID NO: 3, 8, or 9.
  • the RT polypeptide sequence comprises the amino acid sequence of SEQ ID NO: 7, and further comprises a combination of mutations selected from the group consisting of: (i) E69K, L139P, E302R, T306K, W313F, T330P, and N454K; and additionally one or more of M39V, P47L, Q91R, M66L, F155Y, D200N, D200E, H204R, G429S, L435G, L435K, P448A, D449G, H503V, D524N, T542D, E545G, D583N, H594Q, L603W, L603F, E607K, E607G, P627S, H634Y, H638G, A644V, D653H, K658R and L671P; and (ii) E69K, L139P, D
  • the amino acid sequence of the RT polypeptide sequence is: (a) at least 90% identical to SEQ ID NO: 1 or 143; (b) about 90% to about 99.99% identical to SEQ ID NO: 1 or 143, about 92% to about 99.99% identical to SEQ ID NO: 1 or 143, about 93% to about 99.99% identical to SEQ ID NO: 1 or 143, about 94% to about 99.99% identical to SEQ ID NO: 1 or 143, about 95% to about 99.99% identical to SEQ ID NO: 1 or 143, about 96% to about 99.99% identical to SEQ ID NO: 1 or 143, about 97% to about 99.99% identical to SEQ ID NO: 1 or 143, or about 98% to about 4 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC 99.99% identical to SEQ ID NO: 1 or 143; or (c) about
  • the RT polypeptide sequence comprises: (a) an amino acid sequence that is at least 95% identical to SEQ ID NO:1, 7, or 179, and; (b) a combination of mutations indexed to SEQ ID NO:7 or 178 selected from the group consisting of: (i) a combination of variants consisting of a T542D mutation, a D583N mutation, an E607G mutation, an A644V mutation, a D653H mutation, and a K658R mutation; and (ii) a combination of variants consisting of an E545G mutation, a D583N mutation, an H594Q mutation, an L603F mutation, and a S679P mutation.
  • the amino acid sequence of the RT polypeptide sequence comprises E69K, L139P, D200N, E302R, T306K, W313F, T330P, N454K, H503V, D524N, L603W, E607K, and H634Y.
  • the amino acid variations are at any one position or combination thereof as identified in an alignment of SEQ ID NO: 1 or 143 to any one of the RT polypeptide sequences in Table 1 or Table 2.
  • the RT polypeptide sequence comprises: (a) M39V, M66I, Q91R, I347V, and H594Q substitution in SEQ ID NO: 143; or (b) SEQ ID NO: 129 (SOLD 034).
  • the RT polypeptide sequence comprises M39V, T542D, D583N, E607G, A644V, D653H, K658R, and L671P in SEQ ID NO: 143.
  • the RT polypeptide sequence comprises: (a) M39V, T542D, D583N, E607G, A644V, D653H, K658R, L671P in SEQ ID NO: 143; or (b) SEQ ID NO: 111 (SOLD 025).
  • the RT polypeptide sequence comprises T542D, D583N, E607G, A644V, D653H, K658R, E545G, D583N, H594Q, and a L603F in SEQ ID NO: 143.
  • an engineered RT polypeptide comprising: (a) an amino acid sequence that is at least 90%, at least 92%, at least 95%, at least 97%, at least 98%, or at least 99% identical to: (i) an amino acid sequence of an RT disclosed in Table 1, or Table 2; or (ii) SEQ ID NOs: 27-61, 63, 65, 67, 69, 71, 73, 75, 77, 79, 81, 83, 85, 87, 89, 91, 93, 95, 97, 99, 101, 103, 105, 107, 109, 111, 113, 115, 117, 119, 121, 123, 125, 127, 129, 131, 133, 141, 143, 145, 147, 149, 151, 157, 159, 172, or 173; and (b) a DNA binding domain comprising an amino acid selected from the group consisting of SEQ ID NO
  • Another aspect of the present disclosure provides an engineered RT polypeptide comprising: (a) an amino acid sequence of an RT disclosed in Table 1 or Table 2; and (b) an amino acid sequence of DNA binding domain disclosed in Table 1.
  • the engineered RT polypeptide comprises: (a) the amino acid sequence of any one of SEQ ID NO: 174-188; (b) an amino acid sequence having at least 90%, at least 91%, at least 92%, at least 93%, at least 94%, at least 95%, at least 96%, at least 97%, at least 98%, or at least 99% sequence identity to any one of SEQ ID NO: SEQ ID NO: 174-188; or (c) an amino acid sequence having 90%, 91%, 92%, 93%, 94%, 95%, 96%, 97%, 98%, or 99% sequence identity to any one of SEQ ID NO: SEQ ID NO: 174-188.
  • the engineered RT comprises an amino acid sequence that is at least about 90% identical to an amino acid sequence selected from the group consisting of SEQ ID NO: 27-61, 63, 65, 67, 69, 71, 73, 75, 77, 79, 81, 83, 85, 87, 89, 91, 93, 95, 97, 99, 101, 103, 105, 107, 109, 111, 113, 115, 117, 119, 121, 123, 125, 127, 129, 131, 133, 141, 143, 145, 147, 149, 151, 157, 159, 172, and 173.
  • the RT polypeptide is 42B L (SEQ ID NO: 145), 50A+G (SEQ ID NO: 147), SOLD 022 (SEQ ID NO: 105), SOLD 023 (SEQ ID NO: 107), SOLD 025 (SEQ ID NO: 111), SOLD 031 (SEQ ID NO: 123), SOLD 033 (SEQ ID NO: 127), SOLD 034 (SEQ ID NO: 129), SOLD 035 (SEQ ID NO: 131), SOLD 001 (SEQ ID NO: 65), and SOLD 33 VDG (SEQ ID NO: 173), or an RT polypeptide set forth in SEQ ID NO: 143, or SEQ ID NO: 172.
  • the engineered RT comprises at least two DNA binding domains.
  • at least one DNA binding domain is located at the N-terminus of the engineered RT and at least one DNA binding domain is located at the C-terminus of the engineered RT.
  • the at least two DNA binding domains are both located at the C-terminus or N-terminus of the engineered RT.
  • RT reverse transcriptase
  • RT reverse transcriptase
  • RT polypeptide is any one of the RT polypeptides listed in Table 1 or Table 2.
  • the DNA binding domain is a DNA binding protein selected from the group consisting of S. cerevisiae datin (DAT1); high mobility group AT hook 1 (HMGA1), lysine-specific methyltransferase 2a ( KMT2A), Myocyte Enhancer Factor 2C (MEF2C), Heterogeneous Nuclear Ribonucleoprotein D (HNRNPD), Structural Maintenance of Chromosomes 1A (SMC1), Structural Maintenance Of Chromosomes 2 (SMC2), C. elegans tbp-1, D.
  • DAT1 S. cerevisiae datin
  • HMGA1 high mobility group AT hook 1
  • KMT2A lysine-specific methyltransferase 2a
  • MEF2C Myocyte Enhancer Factor 2C
  • HNRNPD Heterogeneous Nuclear Ribonucleoprotein D
  • SMC1 Structural Maintenance of Chromosomes 1A
  • SMC2 Structural Maintenance Of
  • the linker is a glycine-serine (GS) linker selected from the group consisting of (GS)n, (GSGGS)n, (SGGSG)n, (GGGS)n, (GGSG)n, (GGSGG)n, (GSGSG)n, (GSGGG)n, GGGSG)n, and (GSSSG)n, where n represents an integer of at least 1.
  • the linker is GGGS.
  • the linker is SGGSG.
  • the DNA binding domain is a S. cerevisiae datin (DAT1) DNA binding domain or fragment thereof.
  • the DNA binding domain comprises: (a) an amino acid sequence selected from the group consisting of SEQ ID NO: 2, 3, 5, 6, 8, 9, and 11-24; or (b) a nucleic acid sequence of SEQ ID NO: 25.
  • the RT polypeptide is selected from the group consisting of 42B L (SEQ ID NO: 145), 50A+G (SEQ ID NO: 147), SOLD 022 (SEQ ID NO: 105), SOLD 023 (SEQ ID NO: 107), SOLD 025 (SEQ ID NO: 111), SOLD 031 (SEQ ID NO: 123), SOLD 033 7 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC (SEQ ID NO: 127), SOLD 034 (SEQ ID NO: 129), SOLD 035 (SEQ ID NO: 131), SOLD 001 (SEQ ID NO: 65), SOLD 33 VDG (SEQ ID NO:
  • Another aspect of the present disclosure provides a recombinant RT protein as described herein comprising, consisting essentially of, or consisting of SEQ ID NO: 174-188.
  • the engineered RT polypeptide described herein, or the recombinant RT protein described herein further comprises a tag protein selected from the group consisting of an affinity tag, a fluorescent tag, or an expression and/or solubility enhancement tag.
  • the tag is selected from hexahistidine tag (his-tag), small ubiquitin-like modifier tag (SUMO), a VariFlex C-Terminal solubility enhancement tag, a short peptide C- terminal tag, Thioredoxin (Trx) tag, Solubility-enhancer peptide sequences (SET) tag, IgG domain B1 of Protein G (GB1) tag, IgG repeat domain ZZ of Protein A (ZZ) tag, Solubility enhancing Ubiquitous Tag (SNUT tag), Seventeen kilodalton protein (Skp tag), Phage T7 protein kinase (T7PK) tag, E.
  • EspA Monomeric bacteriophage T70.3 protein (Orc protein) (Mocr) tag, E. coli trypsin inhibitor (Ecotin) tag, Calcium-binding protein (CaBP) tag, Stress-responsive arsenate reductase (ArsC) tag, N-terminal fragment of translation initiation factor IF2 (IF2-domain I) tag, N-terminal fragment of translation initiation factor IF2 (Expressivity) tag, Fasciola hepatica 8-kDa antigen tag (Fh8), Glutathione-S-transferase (GST) tag, maltose-binding protein tag (MBP), Flag tag peptide (FLAG), streptavidin binding peptide tag (Strep-II; strep), calmodulin-binding protein tag (CBP), mutated dehalogenase tag (HaloTag), staphylococcal Protein A (Protein A), inte
  • the tag is an affinity tag selected from hexahistidine tag (his-tag), Fasciola hepatica 8-kDa antigen tag (Fh8), Glutathione-S-transferase (GST) tag, maltose-binding protein tag (MBP), Flag tag peptide (FLAG), streptavidin binding peptide tag (Strep-II), calmodulin- binding protein tag (CBP), mutated dehalogenase tag (HaloTag), staphylococcal Protein A 8 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC (Protein A), intein mediated purification with the chitin-binding domain (IMPACT (CBD)), cellulose-binding module (CBM), dockerin domain of Clostridium josui tag (Dock
  • the engineered RT polypeptide or the recombinant RT protein comprises: (a) an hexahistidine tag (his-tag); or (b) an amino acid sequence of SEQ ID NO: 62; or an amino acid sequence having at least 90% sequence identity to the amino acid sequence of SEQ ID NO: 62.
  • the engineered RT polypeptide or the recombinant RT protein comprises a solubility enhancer tag selected from the group consisting of a SUMO tag, a GST tag, a Trx tag, a VariFlex C-Terminal solubility enhancement tag, a short peptide C-terminal tag, an Fh8 tag, MBP tag, SET tag, GB1 tag, ZZ tag, HaloTag, SNUT tag, Skp tag, T7PK tag, EspA tag, Mocr tag, Ecotin tag, CaBO tag, ArsC tag, IF2-domain I tag, Expressivity tag, RpoA, tag, SlyD, tag, Tsf tag, RpoS tag, PotD tag, Crr tag, msyB tag, yigD tag, and rpoD tag.
  • a solubility enhancer tag selected from the group consisting of a SUMO tag, a GST tag, a Trx tag, a VariFlex C-Term
  • the engineered RT polypeptide or the recombinant RT protein comprises: (a) a short peptide C-terminal tag; (b) an amino acid sequence of SEQ ID NO: 193; or (c)an amino acid sequence having at least 90% sequence identity to the amino acid sequence of SEQ ID NO: 193.
  • the tag further comprises: (a) an endoprotein cleavage sequence; (b) a cleavage sequence recognized by an endoprotein selected from the group consisting of alanine carboxypeptidase, Armillaria mellea astacin, bacterial leucyl aminopeptidase, cancer procoagulant, cathepsin B, clostripain, cytosol alanyl aminopeptidase, elastase, endoproteinase Arg-C, enterokinase (EnTK), gastricsin, gelatinase, Gly-X carboxypeptidase, glycyl endopeptidase, human rhinovirus 3C protease, hypodermin C, Iga-specific serine endopeptidase, leucyl aminopeptidase, leucyl endopeptidase, ly
  • the engineered RT polypeptide or the recombinant RT protein described herein exhibits increased template switching (TS) efficiency, increased processivity efficiency, increased binding affinity, increased transcription efficiency, increased chemical tolerance, improved ability to yield mitochondrial unique molecular identity (UMI) counts, improved ability to yield ribosomal unique molecular identity (UMI) counts, longer shelf life, higher strand displacement, higher end-to-end template jumping, or any combination thereof, when compared to an RT polypeptide or a recombinant RT protein lacking a conjugated DNA binding domain.
  • TS template switching
  • UMI mitochondrial unique molecular identity
  • UMI ribosomal unique molecular identity
  • the engineered RT polypeptide or the recombinant RT protein comprises at least two or more of increased template switching (TS) efficiency, increased processivity efficiency, increased binding affinity, increased transcription efficiency, increased chemical tolerance, improved ability to yield mitochondrial unique molecular identity (UMI) counts, longer shelf life, higher strand displacement, higher end-to-end template jumping, or improved ability to yield ribosomal unique molecular identity (UMI) counts, when compared to an RT polypeptide or a recombinant RT protein lacking a conjugated DNA binding domain.
  • TS template switching
  • UMI mitochondrial unique molecular identity
  • the recombinant RT protein or the engineered RT exhibits increased transcript capture during amplification, when compared to an RT polypeptide or a recombinant RT protein lacking a conjugated DNA binding domain.
  • the DNA binding domain enhances the hybridization of a transcript and a primer during a nucleic acid amplification process, when compared to an RT polypeptide or a recombinant RT protein lacking a conjugated DNA binding domain.
  • the primer comprises a poly-dT or a poly(dT)VN sequence and a non-poly(dT) sequence; and the transcript comprises a poly-dA sequence.
  • the DNA binding domain stabilizes the oligo(A)-oligo(T) based transcript- primer complex during a nucleic acid amplification process.
  • the primer is a barcoded molecule.
  • the transcript is a nucleic acid molecule selected from an RNA, a mRNA, or a DNA.
  • One aspect of the present disclosure provides an isolated nucleic acid molecule encoding: (a) an engineered RT polypeptide described herein; or (b) a recombinant RT protein described herein.
  • the nucleic acid molecule comprises a sequence selected from SEQ ID NO: 25, SEQ ID NO: 136, SEQ ID NO:138, SEQ ID NO:140, SEQ ID NO:142, SEQ ID NO:144, SEQ ID NO:146, SEQ ID NO:148, SEQ ID NO:150, SEQ ID NO:152, SEQ ID NO:154, SEQ ID NO:156, SEQ ID NO:158, SEQ ID NO:160, SEQ ID NO:162, SEQ ID NO:164, SEQ ID NO:167, SEQ ID NO:169, or SEQ ID NO: 171; or a nucleic acid sequence of Table 2.
  • One aspect of the present disclosure provides an expression vector comprising any isolated nucleic acid described herein.
  • One aspect of the present disclosure provides a host cell transfected with any expression vector described herein or any isolated nucleic acid described herein.
  • One aspect of the present disclosure provides a composition comprising: (a) any recombinant RT protein described herein; or (b) any engineered RT polypeptide described herein; or (d) any expression vector described herein; or (e) any host cell described herein; and (f) a buffer.
  • One aspect of the present disclosure provides a method for performing a reverse transcription reaction for generating a nucleic acid product from an RNA template comprising contacting under suitable conditions a biological sample or extract thereof with an engineered RT polypeptide described herein, or a recombinant RT protein describe herein.
  • the biological sample or extract thereof comprises a cell, optionally the cell is permeabilized and/or optionally the cell is fixed.
  • the biological sample or extract thereof comprises a cell bead, optionally the cell bead is fixed.
  • the biological sample or extract thereof comprises a nucleus, optionally the nucleus is permeabilized and optionally the nucleus is fixed. In some embodiments, the biological sample or extract thereof comprises (a) a suitable cellular preparation selected from cell populations and/or single cells, or (b) a tissue. [00058] In some embodiments, the biological sample or extract thereof comprises: (a) a cell, a cell bead, a permeabilized cell, a nucleus, where the nucleus is optionally permeabilized, and/or optionally the cell, the cell bead, the permeabilized cell and/or nucleus are fixed; (b) a suitable cellular preparation selected from cell populations and/or single cells; or (c) a tissue.
  • the biological sample or extract thereof comprises cells in suspension, fresh cells, fixed cells, or cells and tissues immobilized on various solid surfaces.
  • the reverse transcription reaction is part of a single cell RNA sequencing assay.
  • the single cell RNA sequencing assay further comprises, prior to the reverse transcription, partitioning the cell, the cell bead, or the nucleus into a partition.
  • the single cell RNA sequencing assay further comprises, after the reverse transcription reaction, hybridizing the nucleic acid product to an oligonucleotide molecule comprising a partition-specific barcode.
  • the reverse transcription reaction is part of a spatial RNA sequencing assay.
  • the engineered RT polypeptide or the recombinant RT protein enhances template switching (TS) efficiency, processivity 12 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC efficiency, binding affinity, transcription efficiency, chemical tolerance, ability to yield mitochondrial unique molecular identity (UMI) counts, ability to yield ribosomal unique molecular identity (UMI) counts, strand displacement, end-to-end template jumping, or any combination thereof, when compared to an RT polypeptide or a recombinant RT protein lacking a conjugated DNA binding domain.
  • TS template switching
  • UMI mitochondrial unique molecular identity
  • UMI ribosomal unique molecular identity
  • the engineered RT polypeptide or the recombinant RT protein enhances at least two or more of template switching (TS) efficiency, processivity efficiency, binding affinity, transcription efficiency, chemical tolerance, ability to yield mitochondrial unique molecular identity (UMI) counts, strand displacement, end-to-end template jumping, or ability to yield ribosomal unique molecular identity (UMI) counts, when compared to an RT polypeptide or a recombinant RT protein lacking a conjugated DNA binding domain.
  • TS template switching
  • UMI mitochondrial unique molecular identity
  • UMI ribosomal unique molecular identity
  • the recombinant RT protein or the engineered RT enhances transcript capture during amplification, when compared to an RT polypeptide or a recombinant RT protein lacking a conjugated DNA binding domain.
  • the DNA binding domain of the recombinant RT protein or the engineered RT enhances the hybridization of a transcript and a primer during the amplification process, when compared to an RT polypeptide or a recombinant RT protein lacking a conjugated DNA binding domain.
  • the engineered RT polypeptide or the recombinant RT protein comprises: (a) a DNA binding domain comprising an amino acid sequence selected from SEQ ID NO: 2, 3, 5, 6, 8, 9, or 11-24; and (b) an amino acid sequence selected from SEQ ID NOs: 27- 61, 63, 65, 67, 69, 71, 73, 75, 77, 79, 81, 83, 85, 87, 89, 91, 93, 95, 97, 99, 101, 103, 105, 107, 109, 111, 113, 115, 117, 119, 121, 123, 125, 127, 129, 131, 133, 141, 143, 145, 147, 149, 151, 157, 159, 172, or 173.
  • the amino acid sequence of the engineered RT polypeptide or the recombinant RT protein comprises an amino acid sequence having at least about 90%, at least about 92%, at least about 95%, at least about 97%, at least about 98%, or at least about 99% sequence identity to SEQ ID NO: 174-188.
  • the engineered RT polypeptide or the recombinant RT protein comprises: (a) M39V, M66I, Q91R, I347V, and H594Q substitution in SEQ ID NO: 143; (b) SEQ ID NO: 129 (SOLD 034); (c) SOLD 001 (SEQ ID NO: 65); or (d) SOLD 33 VDG (SEQ ID NO: 173).
  • the engineered RT polypeptide or the recombinant RT protein comprises M39V, T542D, D583N, E607G, A644V, D653H, K658R, and L671P in SEQ ID NO: 1 or 143.
  • the engineered RT polypeptide or the recombinant RT protein comprises: (a) M39V, T542D, D583N, E607G, A644V, D653H, K658R, L671P in SEQ ID NO: 143; or (b) SEQ ID NO: 111 (SOLD 025).
  • the engineered RT or the recombinant RT protein comprises a M39V, M66I, Q91R, I347V, H594Q in SEQ ID NO: 143.
  • the engineered RT polypeptide or the recombinant RT protein comprises an amino acid sequence that is at least about 90%, at least about 92%, at least about 95%, at least about 97%, at least about 98%, or at least about 99% identical to an amino acid sequence disclosed in Table 1 or Table 2.
  • One aspect of the present disclosure provides a method of using an engineered RT polypeptide described herein, or a recombinant RT protein described herein, the method comprising contacting the engineered RT polypeptide or the recombinant RT protein with a nucleic acid template under suitable conditions to produce a polymerized nucleic acid product, where the nucleic acid template comprises an RNA, a DNA, or a nucleic acid comprising an unnatural nucleotide.
  • the nucleic acid template comprises an RNA.
  • One aspect of the present disclosure provides a nucleic acid extension method comprising: (a) contacting a target nucleic acid molecule with an engineered reverse transcriptase polypeptide or a recombinant RT protein and a plurality of nucleic acid barcoded molecules comprising a barcode sequence, and (b) incubating the target nucleic acid, the engineered RT polypeptide or the recombinant RT protein and barcoded molecules under suitable conditions in which the barcoded molecules are extended by the engineered RT polypeptide or the recombinant RT protein.
  • the engineered RT polypeptide comprises the amino acid sequence of an engineered RT polypeptide described herein, or a recombinant RT protein described herein. 14 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC [00073]
  • the recombinant RT protein or the engineered RT polypeptide exhibits increased transcript capture during amplification.
  • the DNA binding domain enhances the hybridization of a transcript and a primer during a nucleic acid amplification process.
  • the primer comprises a poly-dT or a poly(dT)VN sequence and a non-poly(dT) sequence; and the transcript comprises a poly-dA sequence.
  • the DNA binding domain stabilizes the oligo(A)-oligo(T) based transcript- primer complex during a nucleic acid amplification process.
  • the primer is a barcoded molecule.
  • the recombinant RT protein or the engineered RT polypeptide described herein performs the first strand complementary DNA (cDNA) reaction.
  • the first strand cDNA is amplified using a DNA polymerase to generate a second strand cDNA.
  • One aspect of the present disclosure provides a method of producing an engineered RT polypeptide or recombinant RT protein of the present disclosure, the method comprising providing a composition comprising a cell lysate and/or cellular fraction comprising the engineered RT polypeptide or the recombinant RT protein and subjecting the composition to protein purification steps so as to produce a substantially purified engineered RT enzyme.
  • kits comprising: (a) a recombinant RT protein described herein; or (b) an engineered reverse transcriptase polypeptide described herein; or (c) the isolated nucleic acid described herein; or (d) an expression vector described herein; or (e) a host cell described herein; or (f) a composition described herein; and (g) instructions.
  • FIG.1A shows an exemplary sandwiching process where a first substrate (e.g., a slide), including a biological sample, and a second substrate (e.g., array slide) are brought into proximity with one another.
  • FIG.1B shows a fully formed sandwich configuration creating a chamber formed from the one or more spacers, the first substrate, and the second substrate.
  • FIG.2A shows a perspective view of an exemplary sample handling apparatus in a closed position.
  • FIG.2B shows a perspective view of an exemplary sample handling apparatus in an open position.
  • FIG.3A shows the first substrate angled over (superior to) the second substrate.
  • FIG.3B shows that as the first substrate lowers, and/or as the second substrate rises, the dropped side of the first substrate may contact a drop of reagent medium.
  • FIG.3C shows a full closure of the sandwich between the first substrate and the second substrate with one or more spacers contacting both the first substrate and the second substrate.
  • FIG.4A shows a side view of the angled closure workflow.
  • FIG.4B shows a top view of the angled closure workflow. 16 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC
  • FIG.5 is a schematic diagram showing an example of a barcoded capture probe, as described herein.
  • FIG.6 shows a schematic illustrating a cleavable capture probe.
  • FIG.7 shows exemplary capture domains on capture probes.
  • FIG.8 shows an exemplary arrangement of barcoded features within an array.
  • FIG.9A shows and exemplary workflow for performing a templated capture and producing a ligation product
  • FIG.9B shows an exemplary workflow for capturing a ligation product from FIG.9A on a substrate.
  • FIG.10 is a schematic diagram of an exemplary analyte capture agent.
  • FIG.11 is a schematic diagram depicting an exemplary interaction between a feature- immobilized capture probe 1124 and an analyte capture agent 1126.
  • FIG.12 shows a schematic diagram of a non-limiting embodiment of a generalized capture probe used in spatial transcriptomics and/or single cell transcriptomic analyses, exemplary applications in addition to general reverse transcription reactions where the engineered reverse transcriptase of the disclosure could be used to extend a capture probe using a captured target nucleic acid as a template, thereby generating a cDNA product.
  • FIG.13 provides a schematic of an exemplary capillary electrophoresis (CE) validation assay process used to test the activity of candidate enzymes.
  • CE capillary electrophoresis
  • step 1 5’-end labeled DNA primers were hybridized to RNA templates at room temperature (approx.25°C); and poly rG- labeled template switching oligonucleotides (rG-TSO) were added to the reaction mixture.
  • step 2 the temperature was raised to about 53°C and first strand cDNA was synthesized with the addition of a poly-C tail (tailing).
  • step 3 template switching and TSO extension were performed.
  • step 4 the amplification product was transferred to a Genetic Analyzer for analysis.
  • FIGs.14A-B provide schematics of an exemplary single cell and spatial assay for transcript capture.
  • FIG.14A illustrates a schematic process of the 5’ single cell assay and FIG.
  • FIG. 14B illustrates a schematic process of the 3’ single cell assay and step 1 of Visium.
  • the first step of both assays is the hybridization of an oligo(A) from an mRNA to an oligo(T)) of a primer and 17 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC binding of reverse transcriptase to the annealed primer-template (which contains an oligo(A)- oligo(T) tract).
  • FIGs.15A-B provide schematics of exemplary Visium/ spatial 3’/5’ workflows.
  • FIG. 15A illustrates a schematic of a Visium 3’ workflow demonstrating a polyA capture (box).
  • FIG. 15A illustrates a schematic of a Visium 3’ workflow demonstrating a polyA capture (box).
  • FIGs.16A-B provide a non-limiting embodiment of a nucleic acid sequence (FIG. 16A) and an amino acid sequence (FIG.16B) of an exemplary engineered reverse transcriptase (RT) polypeptide described in the present disclosure (N-Dat_42BL) comprising a DNA binding domain (e.g., DAT; bold and underline) operably linked to a reverse transcriptase (42BL) via a linker (bold).
  • FIGs.17A-B provide a non-limiting embodiment of a nucleic acid sequence (FIG.
  • FIG.17A an amino acid sequence (FIG.17B) of an exemplary engineered reverse transcriptase (RT) polypeptide described in the present disclosure (C-Dat_42BL) comprising a reverse transcriptase (42BL) operably linked to a DNA binding domain (e.g., DAT; bold and underline) via a linker (bold).
  • RT reverse transcriptase
  • C-Dat_42BL exemplary engineered reverse transcriptase polypeptide described in the present disclosure
  • 42BL a reverse transcriptase operably linked to a DNA binding domain (e.g., DAT; bold and underline) via a linker (bold).
  • FIGs.18A-B provide a non-limiting embodiment of a nucleic acid sequence (FIG.
  • FIG.18B an amino acid sequence (FIG.18B) of an exemplary engineered reverse transcriptase (RT) polypeptide described in the present disclosure (N-DAT1full_42BL) comprising a DNA binding domain (e.g., full length DAT1 protein; bold and underline) operably linked to a reverse transcriptase (42BL) via a linker (bold).
  • RT reverse transcriptase
  • N-DAT1full_42BL exemplary engineered reverse transcriptase polypeptide described in the present disclosure
  • FIGs.19A-B provide a non-limiting embodiment of a nucleic acid sequence (FIG.
  • FIG.19B an amino acid sequence of an exemplary engineered reverse transcriptase (RT) polypeptide described in the present disclosure (C-DAT1full_42BL) comprising a reverse transcriptase (42BL) operably linked to a DNA binding domain (e.g., full length DAT1 protein; bold and underline) via a linker (bold).
  • RT reverse transcriptase
  • FIG.20 shows the performance of two engineered RTs described herein (N-DAT 42BL; SEQ ID NO: 175) and C-DAT 42BL (SEQ ID NO: 174)) in a Single Cell 5’ (SC-5’) gene expression assay when compared to two control MMLV variants (SOLD 001 (SEQ ID NO: 65) and SOLD 33 VDG (SEQ ID NO: 175)); and illustrates the superiority of the DAT engineered 18 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC RTs single cell assays over control non-DAT engineered RTs based on the median genes and UMIs per cell at 20k and 50k raw-reads per cell (rrpc).
  • FIG.21 shows the performance of two engineered RTs described herein (N-DAT 42BL; SEQ ID NO: 175) and C-DAT 42BL (SEQ ID NO: 174)) in a Single Cell 5’ (SC-5’) gene expression assay when compared to two control MMLV variants (SOLD 001 (SEQ ID NO: 65) and SOLD 33 VDG (SEQ ID NO: 175)); and illustrates the superiority of the DAT engineered RTs for single cell assays over control non-DAT engineered RTs based on spatial transcriptomics and single cell transcriptomic analyses.
  • FIGs.22A-B provide a bar graph (FIG.22A) and a quantification (FIG.22B) of the relative differences in performance between three reverse transcriptases when compared to a control RT and illustrating that a clear performance gains can be seen for a DAT fusion on either the N-terminal or C-terminal domain.
  • FIGs.22A-B show median genes and UMIs/cell at 50k raw-reads per cell of three reverse transcriptases with and without a DAT fusion domain.
  • FIGs.23A-D provide graphs illustrating the performance of various engineered RTs disclosed herein at maximum normalization depth; and showing median genes (FIG.23A) and UMIs/cell (FIG.23B) at maximum normalized read depth comparing library complexity of three reverse transcriptases with and without the DAT fusion domain.
  • FIGs.23C-D show saturation curves of the median genes (FIG.23C) and counts/cell (FIG.23C) as a function of read depth. At maximum normalized read depth, the benefit of the DAT fusion on each RT backbone can clearly be seen.
  • FIGs.24A-F provide graphs illustrating the differential gene expression of some engineered RTs comprising DAT1 at the N-terminus.
  • FIGs.24A, C, and E feature scatter plots showing gene expression correlation of three reverse transcriptases with and without the DAT fusion domain.
  • FIGs.24B, D, and F feature volcano plots showing the number of differentially expressed genes between three reverse transcriptases with and without the DAT fusion domain.
  • FIGs.25A-F provide graphs illustrating the differential gene expression of some engineered RTs comprising DAT1 at the C-terminus.
  • FIGs.25A, C, and E feature scatter plots showing gene expression correlation comparison of engineered RTs comprising N-Terminal 19 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC DAT Domain and C-Terminal DAT Domain.
  • FIGs.25B 42BL versus 42BL NDAT
  • D and F feature volcano plots comparing the number of differentially expressed genes between various RTs. Gains in performance were still present but were not as significant as a C-terminal DAT fusion.
  • the C-terminal DAT fusion did not perform as well as the N-terminal DAT fusion, but both N-terminal DAT fusion and C-terminal DAT fusion showed performance gains when compared to a control RT variant.
  • the control RT variant was a non-DAT fusion RT comprising the same RT backbone (e.g., RT backbone alone; SEQ ID NO: 145).
  • FIGs.26A-D provide performance comparison of the impact of the DAT domain across three reverse transcriptase backbones.
  • FIGs.26A-B summarize median genes (FIG.26A) and UMIs/cell (FIG.26B) at maximum normalization depth comparing the performance among three MMLV RT variants 42B, 42BL, and 50A+ G backbones with and without a N-terminal DAT fusion domain; and illustrate a clear performance benefit from the DAT domain.
  • FIG.26C shows gene expression correlation.
  • FIG.26D shows differential gene expression. MMLV RT variants without a DAT fusion domain were used as controls.
  • FIG.27 provides a schematic of exemplary Visium/ spatial 3’ workflows highlighting the improvements associated with the engineered RT variant comprising DAT1 at the N- terminus (NDAT1) disclosed herein.
  • NDAT1 RT variant reduced transcript mislocalization and increased target capture. Cleaving oligos after annealing and reducing steric hindrance also enhanced optimization.
  • reverse transcription there was an improved template switching and increased RT efficiency.
  • second strand synthesis there was an improved synthesis efficiency and increased product.
  • FIGs.28A-F provide UMI heat maps showing globally detected gene expression in two replicates of spatial assay performed using a first slide configuration, on a human tonsil tissue, using an NDAT1 MMLV RT variant (FIGs.28C-F) or a control MMLV RT variant (FIGs.28A-B) and three different buffer formulations – a commercially available RT reagent (FIG.28A and FIG.28C), Buffer X.4 (FIG.28D), and Buffer X.5 (FIG.28B, FIG.28E, and 20 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC FIG.28F).
  • FIGs.28A-F show that NDAT1 RT variant improved the sensitivity of the assay while maintaining good spatial resolution when compared to the control RT variant, with all three buffer conditions.
  • FIGs.29A-E provide graphs quantifying the quality, sensitivity, and detection of gene expression under the same six conditions shown in FIGs.28A-F.
  • fraction reads in spots under tissue (FIG.29A), reads mapped confidently to transcriptome (FIG.29B), total genes detected (FIG.29C), GRch38 median UMI counts per spot (30k mapped spot-reads per spot) (FIG.29D), and GRch38 median genes per spot (30k mapped spot-reads per spot) (FIG.29E).
  • the quantification confirmed that NDAT1 RT variant improved the sensitivity of the assay while maintaining good spatial resolution when compared to the control RT variant (FIGs.29D-E).
  • a MMLV RT variant without a DAT fusion domain was used as a control.
  • FIGs.30A-F provide UMI heat maps showing gene expression of RGS3 in a human tonsil tissue analyzed using NDAT1 RT variant (FIGs.30C-F) or a control RT variant (FIGs. 30A-B) and three different buffer formulations - RT reagent (FIG.30A and FIG.30C), Buffer X.4 (FIG.30D), and Buffer X.5 (FIG.30B, FIG.30E, and FIG.30F).
  • a MMLV RT variant without a DAT fusion domain was used as a control.
  • UMI heat maps are shown as a log10(UMI) ranging from cold to hot (0, 0.5, 1.0, 1.5, and 2.0).
  • FIGs.31A-C provide UMI heat maps showing gene expression of KRT5 in two replicates of a spatial assay performed using a first slide configuration on human tonsil tissues analyzed using NDAT1 RT variant (FIGs.31C-F) or a control RT variant (FIGs.31A-B; SEQ ID NO: 1, 142, 143, or 172) and three different buffer formulations - RT reagent (FIG.31A and FIG.31C), Buffer X.4 (FIG.31D), and Buffer X.5 (FIG.31B, FIG.31E, and FIG.31F).
  • a MMLV RT variant without a DAT fusion domain was used as a control.
  • FIGs.32A-E provide UMI heat maps showing globally detected gene expression (FIGs.32A-B) or gene expression of Crgm1 (FIG.32C), KRT5 (FIG.32D), and Rho (FIG.
  • FIGs.33A-D provide UMI heat maps showing globally detected gene expression (FIG.33A) or gene expression of CRGM1 (FIG.33B), KRT5 (FIG.33C), and RHO (FIG. 33D) in a spatial assay performed using a second slide configuration (e.g., Visium high definition (HD) spatial gene expression or next generation spatial gene expression) on Zebrafish (non-human or mouse tissue) analyzed using NDAT1 RT variant (42BL-NDAT1) and RT reagent buffer under sandwich configuration conditions.
  • UMI heat maps are shown as a log normalized per experiment for CRYGM1 (0.00-4.0), KRT5 (0.0-9.0) or RHO (0.0-7.0).
  • the present disclosure leverages 22 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC functional characteristics of DNA binding proteins to engineer RT variants that increase the sensitivity of the RT in single cell applications.
  • the present disclosure provides engineered RT variants (e.g., engineered reverse transcriptase (RT) polypeptide or recombinant RT proteins) to increase transcript capture during single application using an accessory protein that is sequence specific for oligo(A)-oligo(T) tract.
  • engineered RT variants e.g., engineered reverse transcriptase (RT) polypeptide or recombinant RT proteins
  • the accessory protein contemplated by the present disclosure is a DNA binding protein or amino acid motif which is sequence specific to oligo(A)-oligo(T) tract.
  • the accessory protein can function as a “fusion” partner with the reverse transcriptase (e.g., MMLV RT or variant thereof) in single cell 5', single cell 3', and/or spatial assays (e.g., 10X Genomics single cell 5', single cell 3', and/or spatial assays).
  • This strategy enabled improvement of transcript capture, and therefore sensitivity of the assays.
  • DAT1 [000122] Dating (DAT1) or a truncation thereof was identified as a candidate sequence specific DNA binding motif.
  • DAT1 is a yeast protein (e.g., Saccharomyces cerevisiae) that specifically recognizes the minor groove of non-alternating oligo(A)-oligo(T) tracts (e.g., >10 bp oligo(A)- oligo(T) tract).
  • yeast protein e.g., Saccharomyces cerevisiae
  • the sequence specific recognition may be determined by three repeated pentads of G-R-K-P-G (SEQ ID NO: 11).
  • the N-terminal 90 amino acids (D90) and/or the N-terminal 36 amino acids (D36) can bind in a sequence specific manner to oligo(A)-oligo(T) tract.
  • DAT1(D-90) can specifically bind to A-T tracts with Kd of about 3 x 10 -10 M (or 3 x 10 -9 M); and DAT1(D-36) protein can bind to A- T tracts with Kd of 4 x 10 -10 M.
  • DAT1(D-90) can also be more resistant to degradation by bacterial proteases than longer and shorter DAT1 derivatives.
  • DAT1(D-90) The DNA binding activity of DAT1(D-90) can also be resistant to heat (boiling in water bath for 10 min) and chemical treatment (6 M guanidine HCl). DAT1(D-90) can also be highly soluble in physiologic salt and pH conditions. [000123] While DNA binding proteins have been used in combination with a reverse transcriptase as fusion proteins to improve processivity of the reverse transcriptase, sequence specific DNA binding proteins have not been explored. See e.g., Oscorbin et al., FEBS Lett., 594, 4338 (2020).
  • sequence specific DNA binding proteins have not been applied 23 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC to sequencing applications of any type. Fusion proteins with DAT1 have also not been constructed. [000124] As such, engineered RT molecules comprising DAT1 or fragments thereof were investigated as methods to improve the sensitivity of single cell assays. The binding properties of DAT1 was found to be highly tunable. DAT1 was combined with reverse transcriptase (42B or other MMLV variants). Both N-terminal and C-terminal fusions were explored.
  • DAT1(90), first 90 N-terminal amino acids; or DAT1(36), a 36 residue minimal DAT1 binding domain i.e., DAT1(90), first 90 N-terminal amino acids; or DAT1(36), a 36 residue minimal DAT1 binding domain.
  • DAT1 binding domain can assist a reverse transcriptase in binding to primed transcripts (FIGs. 14A-B and 15A-B; box), thereby increasing the assay sensitivity.
  • the DAT1 constructs were optimized for: (1) truncation of DAT1 binding domain; (2) tuning the binding strength of DAT1 to oligo(A)-oligo(T) tract through sequence modification (e.g., removal or alteration of G-R-K- P-G binding pentad); and (3) identification and testing of other sequence-specific DNA binding proteins.
  • the DNA binding domain is from a molecule capable of binding a minor groove of a nucleic acid (e.g., DAT 1).
  • the present disclosure provides recombinant reverse transcriptase (RT) proteins comprising a RT polypeptide, fused to a DNA binding domain, where the RT polypeptide and the DNA binding domain are separated by an amino acid linker, and the DNA binding domain is fused to the N-terminus of the RT polypeptide.
  • RT reverse transcriptase
  • the present disclosure provides recombinant reverse transcriptase (RT) proteins comprising a RT polypeptide, fused to a DNA binding domain, where the RT polypeptide and the DNA binding domain are separated by an amino acid linker, and the DNA binding domain is fused to the C-terminus of the RT polypeptide.
  • compositions comprising the engineered RT or the recombinant RT protein and methods of using the engineered RT or the recombinant RT protein for performing reverse transcription reactions in a variety of applications.
  • An exemplary DNA binding protein is DAT1.
  • a full-length DNA binding protein, truncations or fragments thereof, 24 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC and/or peptide motifs with sequence specific binding can be used to engineer the RT contemplated by the present disclosure.
  • FIG.20 and FIG.21 show that DAT1 in combination with reverse transcriptase MMLV, or MMLV variants thereof (e.g., SOLD 001 or SOLD 033 VDG), either fused at the N- terminus or the C-terminus of the RT improved the assay (GEX) sensitivity even at low sequencing depth.
  • the engineered RT molecules disclosed herein exhibited large change in differential gene expression in single cell assays.
  • the engineered RT molecules comprising DAT1 or variants thereof disclosed herein picked-up up to about 5000 additional genes when compared to a non-DAT1 RT control (e.g., MMLV variant alone).
  • FIGs.20-26 The engineered RT molecules disclosed herein also exhibited increase in median UMI counts per spot and median gene counts per spot in spatial assay.
  • the engineered RT molecules disclosed herein gave decrease in fraction of reads mapped to exons with gain in fraction mapped to introns. See e.g., FIG.21. A performance difference between FPLC and plate purified proteins was also demonstrated. [000128]
  • the engineered RT polypeptides and/or the recombinant RT proteins disclosed herein were tested in spatial platforms using the 3’ workflow shown in FIG.27 under sandwich configuration conditions and using three different buffer formulations described herein.
  • the engineered N-DAT1 RT variants disclosed herein enhanced the quality and the sensitivity of the spatial assay metrics while maintaining good spatial resolution when compared to a control RT lacking the DAT1 domain (also referred to as “control RT”).
  • a control RT in the context of DAT1-RT fusion, refers to a non-DAT1 fusion RT comprising the same RT backbone as the N-DAT1 RT or the C-DAT1 RT.
  • the control RT is a MMLV variant (SEQ ID NOs: 1, 143, or 172) or MMLV variant L (SEQ ID NO: 145).
  • FIGs.28A-F show UMI heat maps showing globally detected gene expression in a spatial assay performed using a first slide configuration, on human tonsil tissue analyzed using NDAT1 RT variant (FIGs.28C-F) compared to a control RT variant (FIGs.28A-B) using three different buffer formulations.
  • FIGs.29A-F show that the engineered RT polypeptides and/or the recombinant RT proteins disclosed herein (e.g., an engineered RT variant comprising DAT-1 at the N-terminus (N-DAT)) also significantly enhanced the quality and the sensitivity metrics of the spatial assays.
  • FIGs.29A-F quantify the quality, sensitivity, and detection of gene expression metrics under the same six conditions shown in FIGs.28A-F. In particular, fraction reads in spots under tissue (FIG.29A) was substantially the same among the 6 conditions tested. However, the standard deviation of the sample comprising the N-DAT RT and RT reagent buffer (N-DAT_RTR) was smaller.
  • the “fraction reads under tissue” refers to the ratio between reads in the area of direct interaction between the array and the tissue over total reads.
  • the fraction reads under tissue showed the diffusion or transcript mislocalization that may have occurred during transcript release from the tissue or transcript capture onto the array.
  • Reads mapped confidently to transcriptome (FIG.29B) were significantly increased in the samples containing the N-DAT_RT when compared to the five other conditions.
  • the total genes detected appeared the same in all conditions tested, except in samples including the control RT variant and using RT reagent buffer (42B_RTR), which showed significantly reduced total genes detected.
  • FIG.29D Analysis of GRch38 median UMI counts per spot (30k mapped spot-reads per spot) (FIG.29D) showed that median UMI counts per spot were significantly enhanced in all samples comprising an N-DAT RT variant when compared to the control RT samples regardless of the buffer used. Unexpectedly, samples comprising the N- DAT_RT variant and Buffer X.4 showed less variability.
  • the GRch38 median genes per spot (30k mapped spot-reads per spot) (FIG.29E) analysis showed that median genes per spot were significantly enhanced in all samples comprising an N-DAT1 RT variant when compared to the control RT samples regardless of the buffer used. Assessment of individual gene expression also showed substantially similar results as globally detected gene expression.
  • FIGs.29E Analysis of GRch38 median UMI counts per spot (30k mapped spot-reads per spot) (FIG.29D) showed that median UMI counts per spot were significantly enhanced in all samples comprising an N-DAT RT variant when compared to the control RT samples regardless
  • FIGs.31A-C show UMI heat maps illustrating the gene expression of KRT5 in a human tonsil 26 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC tissue.
  • spatial assays conducted with an N-DAT RT variant were successful using two different slide configurations: the first slide configuration (e.g., Visium standard definition (SD) spatial gene expression assay) or the second slide configuration (e.g., Visium high definition (HD) or next generation spatial gene expression assay), with a sample from Zebrafish (FIGs.32A-E and FIGs.33A-D).
  • SD Visium standard definition
  • HD Visium high definition
  • FIGs.32A-E and FIGs.33A-D next generation spatial gene expression assay
  • the engineered RT polypeptides and/or the recombinant RT proteins disclosed herein improved the sensitivity of the spatial assay while maintaining good spatial resolution when compared to the control RT variant (FIGs.29D-E). Furthermore, in some metrics, such as reads mapping confidently to the genome, use of the N-DAT RT variant with RT Reagent was particularly superior.
  • the present disclosure demonstrates for the first time that incorporation of DAT1 or any molecule having substantially similar molecular function, in single and spatially assay increased transcript capture and single cell assay sensitivity.
  • the DNA binding domain enhances the enzymatic activity of the engineered reverse transcriptase.
  • the addition of the DNA binding domain can enhance the template switching (TS) efficiency, higher end-to-end template jumping/switching, processivity efficiency, binding affinity, transcription efficiency, chemical tolerance, ability to yield mitochondrial unique molecular identifier (UMI) counts, ability to yield ribosomal unique molecular identifier (UMI) counts, shelf life, higher strand displacement, increased thermostability, improved thermoreactivity, and any combination thereof, for the engineered (i.e., recombinant) reverse transcriptase when compared to a control RT (SEQ ID NO: 1, 143, 145, or 172), WT MMLV, or known MMLV variants.
  • TS template switching
  • UMI mitochondrial unique molecular identifier
  • UMI ribosomal unique molecular identifier
  • TS efficiency Small RNAs ( ⁇ 200 nucleotides) are for the most part non-coding regulatory elements and play a key role in gene expression. Small RNAs regulate gene expression in plants, animals, and many fungi—including several roles in development, proliferation, differentiation, immune reaction, apoptosis, tumorigenesis and adaptation to stress. Given their importance in regulation, miRNAs are candidates as biomarkers for several human diseases. Thus, developing accurate and reproducible ways to study these and other small RNAs is necessary to further decipher their biological consequences. [000135] The main sources of bias in a typical library preparation workflow are the enzymatic ligations that introduce 5′ and 3′ sequencing adaptors to single-stranded templates.
  • Template 27 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC switching permits ligation-free incorporation of the 5′ adapter during reverse transcription.
  • Template switching-based methods depend upon the natural tendency of MMLV-type reverse transcriptases to add nontemplated nucleotides at the 3′ end of the emerging cDNA strand. These nontemplated additions serve as an anchoring unit for annealing complementary nucleotides in a provided template switching oligonucleotide (TSO); upon reaching the cDNA-TSO cross- junction, the reverse transcriptase effectively switches templates, continuing cDNA synthesis out of the TSO sequence.
  • TSO template switching oligonucleotide
  • End-to-end template jumping or switching refers to the ability of a reverse transcriptase to template-switch from the 5’ end of one template to the 3’ end of another. Improved end-to-end template jumping or switching can result in an improved process efficiency.
  • the engineered reverse transcriptase described herein exhibiting improved or higher end-to-end template jumping or switching, is highly desirable.
  • DNA binding affinity To initiate reverse transcription, reverse transcriptases require a short DNA oligonucleotide called a primer to bind to its complementary sequences on the RNA template and serve as a starting point for synthesis of a new strand. Improved binding affinity results in a more efficient process, particularly when limited amounts of RNA are available. Thus, the engineered reverse transcriptase described herein, exhibiting improved DNA binding affinity, is highly desirable.
  • Transcription efficiency The RNA-to-cDNA conversion step in transcriptomics experiments is widely recognized as inefficient and variable.
  • Transcriptomics measurements almost invariably include a reverse transcription (RT) step, where RNA transcripts are used as templates to generate cDNA transcripts for quantification.
  • RT reverse transcription
  • RNA transcripts are used as templates to generate cDNA transcripts for quantification.
  • the engineered reverse transcriptase described herein exhibiting improved transcription efficiency, is highly desirable.
  • Reverse transcriptases function in an environment that may include processing chemicals, such as cell fixation chemicals or processing reagents, which can negatively impact the function and activity of the enzyme. Thus, the engineered reverse transcriptase described herein, exhibiting improved chemical tolerance, is highly desirable.
  • Ability to yield mitochondrial and/or ribosomal unique molecular identifier (UMI) counts Unique molecular identifier (UMI) counting is a gene expression quantification scheme used in single-cell RNA-sequencing (scRNA-seq) analysis. Single-cell RNA-sequencing (scRNA-seq) technology provides transcriptome profiles of individual cells, enabling the dissection of the heterogeneity of different cell populations and tissues.
  • scRNA-seq protocols The paucity of starting material for reverse transcription remains an inherent limitation of scRNA-seq protocols and contributes to the relatively low rate at which messenger RNA (mRNA) molecules in individual cells are converted to cDNA molecules that can be captured and sequenced.
  • mRNA messenger RNA
  • mRNA-seq protocols employ an additional step in which individual transcripts are barcoded with unique molecular identifiers (UMIs) before amplification, resulting in a more accurate quantification of the transcript count.
  • UMIs incorporate a unique barcode onto each molecule within a given sample library.
  • variant alleles present in the original sample can be distinguished from errors introduced during library preparation, target enrichment, or sequencing.
  • the engineered reverse transcriptase described herein exhibiting an improved ability to yield mitochondrial and/or ribosomal UMI counts, is highly desirable.
  • Shelf life and/or stability In another aspect of the disclosure, the engineered reverse transcriptase described herein, exhibit improved stability and/or shelf life. A longer period of stability, and/or shelf life, is desirable as it can result in more efficient processes.
  • Strand displacement is the process through which two strands with partial or full complementarity hybridize to each other, displacing one or more pre- hybridized strands in the process.
  • Reverse transcriptase first transcribes a complementary strand 29 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC of DNA to make an RNA:DNA hybrid.
  • reverse transcriptase or RNase H degrades the RNA strand of the hybrid.
  • the single-stranded DNA is then used as a template for synthesizing double-stranded DNA (cDNA).
  • RT reverse transcriptase catalyzes the conversion of RNA into an integration-competent double-stranded DNA, with a variety of enzymatic activities that include the ability to displace a non-template strand concomitantly with polymerization.
  • RT are capable of efficiently unwinding duplexes in the template during polymerization.
  • This strand displacement synthesis activity by RT is required for the polymerization on the highly structured RNA and the removal of RNA fragments which cannot be cleaved by the enzymes’ RNase H activity.
  • strand displacement synthesis on a DNA duplex is particularly important to complete the plus- and minus-strands by polymerizing on the long terminal repeats.
  • any of the engineered RT enzymes of the present disclosure including without limitation any of the enzymes comprising the amino acid sequence and/or nucleic acid sequences shown in Table 1 or Table 2 could be analyzed in any suitable assay, including without limitation the assays described herein. Assays include without limitation 5’ gene expression analyses, with or without VDJ analysis, 3’ gene expression analysis, epigenetic analysis, or multiomic analyses.
  • experiments are carried out as found in the manufacturer’s instructions for the Chromium Single Cell 5’ Gene Expression Assay kit (10X Genomics); Chromium Single Cell 3’ Gene Expression Assay kit (10X Genomics), including any of multiomic extensions or applications.
  • SPATIAL ANALYSIS METHODS [000144] Spatial analysis methodologies described herein can provide a vast amount of analyte and/or expression data for a variety of analytes within a biological sample at high spatial resolution, while retaining native spatial context.
  • Spatial analysis methods can include, e.g., the use of a capture probe including a spatial barcode (e.g., a nucleic acid sequence that provides information as to the location or position of an analyte within a cell or a tissue sample (e.g., mammalian cell or a mammalian tissue sample) and a capture domain that is capable of binding to an analyte (e.g., a protein and/or a nucleic acid) produced by and/or present in a cell.
  • a spatial barcode e.g., a nucleic acid sequence that provides information as to the location or position of an analyte within a cell or a tissue sample
  • a capture domain that is capable of binding to an analyte (e.g., a protein and/or a nucleic acid) produced by and/or present in a cell.
  • Spatial 30 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC analysis methods and compositions can also include the use of a capture probe having a capture domain that captures an intermediate agent for indirect detection of an analyte.
  • the intermediate agent can include a nucleic acid sequence (e.g., a barcode) associated with the intermediate agent. Detection of the intermediate agent is therefore indicative of the analyte in the cell or tissue sample.
  • a nucleic acid sequence e.g., a barcode
  • a “barcode” is a label, or identifier, that conveys or is capable of conveying information (e.g., information about an analyte in a sample, a bead, and/or a capture probe).
  • a barcode can be part of an analyte, or independent of an analyte.
  • a barcode can be attached to an analyte.
  • a particular barcode can be unique relative to other barcodes.
  • an “analyte” can include any biological substance, structure, moiety, or component to be analyzed.
  • the term “target” can similarly refer to an analyte of interest. 31 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC [000147]
  • Analytes can be broadly classified into one of two groups: nucleic acid analytes, and non-nucleic acid analytes.
  • non-nucleic acid analytes include, but are not limited to, lipids, carbohydrates, peptides, proteins, glycoproteins (N-linked or O-linked), lipoproteins, phosphoproteins, specific phosphorylated or acetylated variants of proteins, amidation variants of proteins, hydroxylation variants of proteins, methylation variants of proteins, ubiquitylation variants of proteins, sulfation variants of proteins, viral proteins (e.g., viral capsid, viral envelope, viral coat, viral accessory, viral glycoproteins, viral spike, etc.), extracellular and intracellular proteins, antibodies, and antigen binding fragments.
  • viral proteins e.g., viral capsid, viral envelope, viral coat, viral accessory, viral glycoproteins, viral spike, etc.
  • the analyte(s) can be localized to subcellular location(s), including, for example, organelles, e.g., mitochondria, Golgi apparatus, endoplasmic reticulum, chloroplasts, endocytic vesicles, exocytic vesicles, vacuoles, lysosomes, etc.
  • organelles e.g., mitochondria, Golgi apparatus, endoplasmic reticulum, chloroplasts, endocytic vesicles, exocytic vesicles, vacuoles, lysosomes, etc.
  • analyte(s) can be peptides or proteins, including without limitation antibodies and enzymes. Additional examples of analytes can be found in Section (I)(c) of PCT Publication No. WO2020/176788 and/or U.S. Patent Application Publication No.2020/0277663.
  • an analyte can be detected indirectly, such as through detection of an intermediate agent, for example, a ligation product or an analyte capture agent (e.g., an oligonucleotide-conjugated antibody), such as those described herein.
  • an intermediate agent for example, a ligation product or an analyte capture agent (e.g., an oligonucleotide-conjugated antibody), such as those described herein.
  • an intermediate agent for example, a ligation product or an analyte capture agent (e.g., an oligonucleotide-conjugated antibody), such as those described herein.
  • an intermediate agent for example, a ligation product or an analyte capture agent (e.g., an oligonucleotide-conjugated antibody), such as those described herein.
  • an intermediate agent for example, a ligation product or an analyte capture agent (e.g., an
  • a tissue microarray contains multiple representative tissue samples – which can be from different tissues or organisms – assembled on a single histologic slide.
  • the TMA can therefore allow for high throughput analysis of multiple specimens at the same time.
  • Tissue microarrays are paraffin blocks produced by extracting cylindrical tissue cores from different paraffin donor blocks and re-embedding these into a single recipient (microarray) block at defined array coordinates.
  • the biological sample as used herein can be any suitable biological sample described herein or known in the art.
  • the biological sample is a tissue.
  • the tissue sample is a solid tissue sample.
  • the biological sample is a tissue section.
  • the tissue is flash-frozen and sectioned.
  • any 32 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC suitable method described herein or known in the art can be used to flash-freeze and section the tissue sample.
  • the biological sample e.g., the tissue
  • the biological sample is flash-frozen using liquid nitrogen before sectioning.
  • the biological sample e.g., a tissue sample
  • nitrogen e.g., liquid nitrogen
  • isopentane e.g., or hexane.
  • the biological sample, e.g., the tissue is embedded in a matrix e.g., optimal cutting temperature (OCT) compound to facilitate sectioning.
  • OCT optimal cutting temperature
  • OCT compound is a formulation of clear, water-soluble glycols and resins, providing a solid matrix to encapsulate biological (e.g., tissue) specimens.
  • the sectioning is performed using cryosectioning.
  • the methods further comprise a thawing step, after the cryosectioning.
  • the biological sample can be from a mammal. In some instances, the biological sample is from a human, mouse, or rat.
  • the biological sample can be obtained from non-mammalian organisms (e.g., a plants, an insect, an arachnid, a nematode (e.g., Caenorhabditis elegans), a fungi, an amphibian, or a fish (e.g., zebrafish)).
  • a biological sample can be obtained from a prokaryote such as a bacterium, e.g., Escherichia coli, Staphylococci or Mycoplasma pneumoniae; an archaea; a virus such as Hepatitis C virus or human immunodeficiency virus; or a viroid.
  • a biological sample can be obtained from a eukaryote, such as a patient derived organoid (PDO) or patient derived xenograft (PDX).
  • the biological sample can include organoids, a miniaturized and simplified version of an organ produced in vitro in three dimensions that shows realistic micro-anatomy.
  • Organoids can be generated from one or more cells from a tissue, embryonic stem cells, and/or induced pluripotent stem cells, which can self-organize in three-dimensional culture owing to their self-renewal and differentiation capacities.
  • an organoid is a cerebral organoid, an intestinal organoid, a stomach organoid, a lingual organoid, a thyroid organoid, a thymic organoid, a testicular organoid, a hepatic organoid, a pancreatic organoid, an epithelial organoid, a lung organoid, a kidney organoid, a gastruloid, a cardiac organoid, or a retinal organoid.
  • Subjects from which biological samples can be obtained can be healthy or asymptomatic individuals, individuals that have or are suspected of having a disease (e.g., cancer) or a pre-disposition to a disease, and/or individuals that are in need of therapy or suspected of needing therapy.
  • a disease e.g., cancer
  • a pre-disposition to a disease e.g., cancer
  • 10X Genomics Ref.: 100-165501PC Biological samples can be derived from a homogeneous culture or population of the subjects or organisms mentioned herein or alternatively from a collection of several different organisms, for example, in a community or ecosystem.
  • Biological samples can include one or more diseased cells.
  • a diseased cell can have altered metabolic properties, gene expression, protein expression, and/or morphologic features. Examples of diseases include inflammatory disorders, metabolic disorders, nervous system disorders, and cancer. Cancer cells can be derived from solid tumors, hematological malignancies, cell lines, or obtained as circulating tumor cells.
  • the biological sample e.g., the tissue sample
  • the biological sample is fixed in a fixative including alcohol, for example methanol. In some embodiments, instead of methanol, acetone, or an acetone-methanol mixture can be used. In some embodiments, the fixation is performed after sectioning. In some instances, the biological sample is not fixed with paraformaldehyde (PFA).
  • PFA paraformaldehyde
  • the biological sample when the biological sample is fixed with a fixative including an alcohol (e.g., methanol or acetone-methanol mixture), it is not decrosslinked afterward.
  • the biological sample is fixed with a fixative including an alcohol (e.g., methanol or an acetone-methanol mixture) after freezing and/or sectioning.
  • the biological sample is flash-frozen, and then the biological sample is sectioned and fixed (e.g., using methanol, acetone, or an acetone-methanol mixture).
  • methanol, acetone, or an acetone-methanol mixture when methanol, acetone, or an acetone-methanol mixture is used to fix the biological sample, the sample is not decrosslinked at a later step.
  • the biological sample is frozen (e.g., flash frozen using liquid nitrogen and embedded in OCT) followed by sectioning and alcohol (e.g., methanol, acetone-methanol) fixation or acetone fixation
  • fresh frozen e.g., acetone and/or alcohol (e.g., methanol, acetone-methanol)
  • fixation of the biological sample e.g., using acetone and/or alcohol (e.g., methanol, acetone-methanol) is performed while the sample is mounted on a substrate (e.g., glass slide, such as a positively charged glass slide).
  • the biological sample e.g., the tissue sample
  • the biological sample is fixed e.g., immediately after being harvested from a subject.
  • the fixative is preferably an aldehyde fixative, such as paraformaldehyde (PFA) or formalin.
  • the fixative induces crosslinks within the biological sample.
  • the biological sample is dehydrated via sucrose gradient.
  • the fixed biological sample is treated with a sucrose gradient and then embedded in a matrix e.g., OCT compound.
  • the fixed biological sample is not treated with a sucrose gradient, but rather is embedded in a matrix e.g., OCT compound after fixation.
  • a fixed frozen tissue sample when a fixed frozen tissue sample is treated with a sucrose gradient, it can be rehydrated with an ethanol gradient.
  • the PFA or formalin fixed biological sample which can be optionally dehydrated via sucrose gradient and/or embedded in OCT compound, is then frozen e.g., for storage or shipment.
  • the biological sample is referred to as “fixed frozen”.
  • a fixed frozen biological sample is not treated with methanol.
  • a fixed frozen biological sample is not paraffin embedded.
  • a fixed frozen biological sample is not deparaffinized.
  • a fixed frozen biological sample is rehydrated in an ethanol gradient.
  • the biological sample e.g., a fixed frozen tissue sample
  • Citrate buffer can be used for antigen retrieval to decrosslink antigens and fixation medium in the biological sample.
  • any suitable decrosslinking agent can be used in addition to or alternatively to citrate buffer.
  • the biological sample e.g., a fixed frozen tissue sample
  • the biological sample can further be stained, imaged, and/or destained.
  • a fresh frozen tissue sample or fixed frozen tissue sample is stained (e.g., via eosin and/or hematoxylin), imaged, destained (e.g., via HCl), or a combination thereof.
  • a fresh frozen tissue sample is fixed in methanol, it is treated with isopropanol prior to being stained (e.g., via eosin and/or hematoxylin), imaged, destained (e.g., via HCl), or a combination thereof.
  • a fixed frozen tissue sample when a fixed frozen tissue sample is treated with a sucrose gradient, it can be rehydrated with an ethanol gradient before being stained, (e.g., via eosin and/or hematoxylin), imaged, destained (e.g., via HCl), decrosslinked (e.g., via TE buffer or citrate buffer), or a combination thereof.
  • the biological sample can undergo further fixation (e.g., while mounted on a substrate), stained, imaged, and/or destained.
  • a fixed frozen biological sample may 35 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC be subject to an additional fixing step (e.g., using PFA) before optional ethanol rehydration, staining, imaging, and/or destaining.
  • the biological sample can be fixed using PAXgene ® .
  • the biological sample can be fixed using PAXgene ® in addition, or alternatively to, a fixative disclosed herein or known in the art (e.g., alcohol, acetone, acetone-alcohol, formalin, paraformaldehyde).
  • PAXgene ® is a non-cross-linking mixture of different alcohols, acid and a soluble organic compound that preserves morphology and bio-molecules. It is a two-reagent fixative system in which tissue is firstly fixed in a solution containing methanol and acetic acid then stabilized in a solution containing ethanol. See, Ergin B. et al., J Proteome Res.2010 Oct 1;9(10):5188-96; Kap M. et al., PLoS One.; 6(11):e27704 (2011); and Mathieson W.
  • the fixative is PAXgene ® .
  • a fresh frozen tissue sample is fixed with PAXgene ® .
  • a fixed frozen tissue sample is fixed with PAXgene ® .
  • the biological sample e.g., the tissue sample is fixed, for example in methanol, acetone, acetone-methanol, PFA, PAXgene ® or is formalin-fixed and paraffin-embedded (FFPE).
  • the biological sample comprises intact cells.
  • the biological sample is a cell pellet, e.g., a fixed cell pellet, e.g., an FFPE cell pellet. FFPE samples are used in some instances in the RTL methods disclosed herein.
  • RNA integrity of fixed (e.g., FFPE) samples can be lower than a fresh sample, thereby making it more difficult to capture RNA directly, e.g., by capture of a common sequence such as a poly(A) tail of an mRNA molecule.
  • RTL probes can be utilized to beneficially improve capture and spatial analysis of fixed samples.
  • the biological sample e.g., tissue sample, can be stained, and imaged prior, during, and/or after each step of the methods described herein.
  • tissue sample can be obtained from any suitable location in a tissue or organ of a subject, e.g., a human subject.
  • the sample is a mouse sample.
  • the sample is a human sample.
  • the sample can be derived from skin, brain, breast, lung, liver, kidney, prostate, tonsil, thymus, testes, bone, lymph node, ovary, eye, heart, or spleen.
  • the sample is a human or mouse breast tissue sample.
  • the sample is a human or mouse brain tissue sample.
  • the sample is a human or mouse lung tissue sample.
  • the sample is a human or mouse tonsil tissue sample.
  • the sample is a human or mouse liver tissue sample.
  • the sample is a human or mouse bone, skin, kidney, thymus, testes, or prostate tissue sample.
  • the tissue sample is derived from normal or diseased tissue.
  • the sample is an embryo sample.
  • the embryo sample can be a non-human embryo sample.
  • the sample is a mouse embryo sample.
  • stains include histological stains (e.g., hematoxylin and/or eosin) and immunological stains (e.g., fluorescent stains).
  • the biological sample is imaged.
  • the biological sample is visualized or imaged using bright field microscopy.
  • the biological sample is visualized or imaged using fluorescence microscopy. Additional methods of visualization and imaging are known in the art. Non-limiting examples of visualization and imaging include expansion microscopy, bright field microscopy, dark field microscopy, phase 37 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC contrast microscopy, electron microscopy, fluorescence microscopy, reflection microscopy, interference microscopy and confocal microscopy.
  • the sample is stained and imaged prior to adding the primer to the biological sample.
  • the method includes staining the biological sample. In some embodiments, the staining includes the use of hematoxylin and eosin.
  • a biological sample can be stained using any number of biological stains, including but not limited to, acridine orange, Bismarck brown, carmine, coomassie blue, cresyl violet, DAPI, eosin, ethidium bromide, acid fuchsine, hematoxylin, Hoechst stains, iodine, methyl green, methylene blue, neutral red, Nile blue, Nile red, osmium tetroxide, propidium iodide, rhodamine, or safranin.
  • biological stains including but not limited to, acridine orange, Bismarck brown, carmine, coomassie blue, cresyl violet, DAPI, eosin, ethidium bromide, acid fuchsine, hematoxylin, Hoechst stains, iodine, methyl green, methylene blue, neutral red, Nile blue, Nile red, osm
  • the biological sample can be stained using known staining techniques, including Can-Grunwald, Giemsa, hematoxylin and eosin (H&E), Jenner’s, Leishman, Masson’s trichrome, Papanicolaou, Romanowsky, silver, Sudan, Wright’s, and/or Periodic Acid Schiff (PAS) staining techniques.
  • PAS staining is typically performed after formalin or acetone fixation.
  • the staining includes the use of a detectable label selected from the group consisting of a radioisotope, a fluorophore, a chemiluminescent compound, a bioluminescent compound, or a combination thereof.
  • a biological sample is permeabilized with one or more permeabilization reagents.
  • permeabilization of a biological sample can facilitate analyte capture.
  • Exemplary permeabilization agents and conditions are described in Section (I)(d)(ii)(13) or the Exemplary Embodiments Section of PCT Publication No. WO2020/176788 and/or U.S. Patent Application Publication No.2020/0277663.
  • the method includes a step of permeabilizing the biological sample.
  • the biological sample can be permeabilized to facilitate transfer of the extension products to the capture probes on the array.
  • the permeabilizing includes the use of an organic solvent (e.g., acetone, ethanol, and methanol), a detergent (e.g., saponin, Triton X-100TM, Tween-20TM, or sodium dodecyl sulfate (SDS)), an enzyme (an endopeptidase, an exopeptidase, a protease), or combinations thereof.
  • an organic solvent e.g., acetone, ethanol, and methanol
  • a detergent e.g., saponin, Triton X-100TM, Tween-20TM, or sodium dodecyl sulfate (SDS)
  • an enzyme an endopeptidase, an exopeptidase, a protease
  • the permeabilizing includes the use of an endopeptidase, a protease, SDS, polyethylene glycol tert-octylphenyl ether, polysorbate 80, and polysorbate 20, N-lauroylsarcosine sodium salt solution, saponin, 38 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC Triton X-100TM, Tween-20TM, or combinations thereof.
  • the endopeptidase is pepsin.
  • the endopeptidase is Proteinase K.
  • Array-based spatial analysis methods involve the transfer of one or more analytes from a biological sample to an array of features on a substrate, where each feature is associated with a unique spatial location on the array. Subsequent analysis of the transferred analytes includes determining the identity of the analytes and the spatial location of the analytes within the biological sample.
  • a “capture probe” refers to any molecule capable of capturing (directly or indirectly) and/or labelling an analyte (e.g., an analyte of interest) in a biological sample.
  • the capture probe is a nucleic acid or a polypeptide.
  • the capture probe includes a barcode (e.g., a spatial barcode and/or a unique molecular identifier (UMI)) and a capture domain).
  • UMI unique molecular identifier
  • the capture probe includes a homopolymer sequence, such as a poly(T) sequence.
  • a capture probe can include a cleavage domain and/or a functional domain (e.g., a primer-binding site, such as for next- generation sequencing (NGS)).
  • NGS next- generation sequencing
  • the biological sample is mounted on a first substrate and the substrate comprising the array of capture probes is a second substrate.
  • one or more analytes or analyte derivatives e.g., intermediate agents; e.g., ligation products
  • the release and migration of the analytes or analyte derivatives to the second substrate comprising the array of capture probes occurs in a manner that preserves the original spatial context of the analytes in the biological sample.
  • FIG.1A shows an exemplary sandwiching process 100 where a first substrate (e.g., slide 103), including a biological sample 102, and a second substrate (e.g., array slide 104 including an array having spatially barcoded capture probes 106) are brought into proximity with one another.
  • a first substrate e.g., slide 103
  • a second substrate e.g., array slide 104 including an array having spatially barcoded capture probes 106
  • a liquid reagent drop (e.g., permeabilization solution 105) is introduced on the second substrate in proximity to the capture probes 106 and in between the biological sample 102 and the second substrate (e.g., slide 104 including an array having spatially barcoded capture probes 106).
  • the permeabilization solution 105 may release analytes or analyte derivatives (e.g., intermediate agents; e.g., ligation products) that can be captured by the capture probes of the array 106.
  • the first substrate is aligned with the second substrate, such that at least a portion of the biological sample is aligned with at least a portion of the capture probes (e.g., aligned in a sandwich configuration).
  • the second substrate e.g., array slide 104 is in an inferior position to the first substrate (e.g., slide 103).
  • the first substrate e.g., slide 103
  • the second substrate e.g., slide 104
  • a reagent medium 105 within a gap between the first substrate (e.g., slide 103) and the second substrate (e.g., slide 104) creates a liquid interface between the two substrates.
  • the reagent medium may be a permeabilization solution which permeabilizes and/or digests the biological sample 102.
  • the reagent medium is not a permeabilization solution.
  • analytes e.g., mRNA transcripts
  • analyte derivatives e.g., intermediate agents; e.g., ligation products
  • release from the biological sample and actively or passively migrate (e.g., diffuse) across the gap toward the capture probes on the array 106.
  • migration of the analyte or analyte derivative (e.g., intermediate agent; e.g., ligation product) from the biological sample is performed actively (e.g., electrophoretic, by applying an electric field to promote migration).
  • electrophoretic by applying an electric field to promote migration.
  • one or more spacers 110 may be positioned between the first substrate (e.g., slide 103) and the second substrate (e.g., array slide 104 including spatially barcoded capture probes 106).
  • the one or more spacers 110 may be configured to maintain a separation distance between the first substrate and the second substrate. While the one or more spacers 110 is shown as disposed on the second substrate, the spacer may additionally or alternatively be disposed on the first substrate.
  • the one or more spacers 110 is configured to maintain a separation distance between first and second substrates that is between about 2 microns and 1 mm (e.g., between about 2 microns and 800 microns, between about 2 microns and 700 microns, between about 2 microns and 600 microns, between about 2 microns and 500 microns, between about 2 microns and 400 microns, between about 2 microns and 300 microns, between about 2 microns and 200 microns, between about 2 microns and 100 microns, between about 2 microns and 25 microns, or between about 2 microns and 10 microns), measured in a direction orthogonal to the surface of first substrate that supports the biological sample.
  • a separation distance between first and second substrates that is between about 2 microns and 1 mm (e.g., between about 2 microns and 800 microns, between about 2 microns and 700 microns, between about 2 microns and 600 microns, between about 2 microns and
  • the separation distance is about 2, 3, 4, 5, 6, 7, 8, 9, 10, 11, 12, 13, 14, 15, 16, 17, 18, 19, 20, 21, 22, 23, 24, or 25 microns. In some embodiments, the separation distance is less than 50 microns. In some embodiments, the separation distance is less than 25 microns. In some embodiments, the separation distance is less than 20 microns. The separation distance may include a distance of at least 2 ⁇ m.
  • FIG.1B shows a fully formed sandwich configuration 125 creating a chamber 150 formed from the one or more spacers 110, the first substrate (e.g., the slide 103), and the second substrate (e.g., the slide 104 including an array 106 having spatially barcoded capture probes) in accordance with some example implementations.
  • the liquid reagent e.g., the permeabilization solution 105 fills the volume of the chamber 150 and may create a permeabilization buffer that allows analytes (e.g., mRNA transcripts and/or other molecules) or analyte derivatives (e.g., intermediate agents; e.g., ligation products) to diffuse from the biological sample 102 toward the capture probes of the second substrate (e.g., slide 104).
  • analytes e.g., mRNA transcripts and/or other molecules
  • analyte derivatives e.g., intermediate agents; e.g., ligation products
  • flow of the permeabilization buffer may deflect transcripts and/or molecules from the biological sample 102 and may affect diffusive transfer of analytes or analyte derivatives (e.g., intermediate agents; e.g., ligation products) for spatial analysis.
  • analytes or analyte derivatives e.g., intermediate agents; e.g., ligation products
  • a partially or 41 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC fully sealed chamber 150 resulting from the one or more spacers 110, the first substrate, and the second substrate may reduce or prevent flow from undesirable convective movement of transcripts and/or molecules over the diffusive transfer from the biological sample 102 to the capture probes.
  • the sandwiching process methods described above can be implemented using a variety of hardware components.
  • the sandwiching process methods can be implemented using a sample holder (also referred to herein as a support device, a sample handling apparatus, and an array alignment device). Further details on support devices, sample holders, sample handling apparatuses, or systems for implementing a sandwiching process are described in, e.g., US. Patent Application Pub. No.2021/0189475, and PCT Publ. No. WO 2022/061152 A2, each of which are incorporated by reference in their entirety.
  • the sample holder can include a first member including a first retaining mechanism configured to retain a first substrate comprising a biological sample.
  • the first retaining mechanism can be configured to retain the first substrate disposed in a first plane.
  • the sample holder can further include a second member including a second retaining mechanism configured to retain a second substrate disposed in a second plane.
  • the sample holder can further include an alignment mechanism connected to one or both of the first member and the second member.
  • the alignment mechanism can be configured to align the first and second members along the first plane and/or the second plane such that the sample contacts at least a portion of the reagent medium when the first and second members are aligned and within a threshold distance along an axis orthogonal to the second plane.
  • the adjustment mechanism may be configured to move the second member along the axis orthogonal to the second plane and/or move the first member along an axis orthogonal to the first plane.
  • the adjustment mechanism includes a linear actuator.
  • the linear actuator is configured to move the second member along an axis orthogonal to the plane of the first member and/or the second member.
  • the linear actuator is configured to move the first member along an axis orthogonal to the plane of the first member and/or the second member.
  • the linear actuator is configured to move the first member, the second member, or both the first member and the second member at a velocity of at least 0.1 mm/sec.
  • FIG.2A is a perspective view of an example sample handling apparatus 200 in a closed position in accordance with some example implementations.
  • the sample handling apparatus 200 includes a first member 204, a second member 210, optionally an image capture device 220, a first substrate 206, optionally a hinge 215, and optionally a mirror 216.
  • FIG.2B is a perspective view of the example sample handling apparatus 200 in an open position in accordance with some example implementations.
  • the sample handling apparatus 200 includes one or more first retaining mechanisms 208 configured to retain one or more first substrates 206.
  • the first member 204 is configured to retain two first substrates 206, however the first member 204 may be configured to retain more or fewer first substrates 206.
  • the first substrate 206 and/or the second substrate 212 may be loaded and positioned within the sample handling apparatus 200 such as within the first member 204 and the second member 210, respectively.
  • the hinge 215 may allow the first member 204 to close over the second member 210 and form a sandwich configuration.
  • an adjustment mechanism of the sample handling apparatus 200 may actuate the first member 204 and/or the second member 210 to form the sandwich configuration for the permeabilization step (e.g., bringing the first substrate 206 and the second substrate 212 closer to each other and within a threshold distance for the sandwich configuration).
  • the adjustment mechanism may be configured to control a speed, an angle, a force, or the like of the sandwich configuration.
  • the biological sample (e.g., sample 102 from FIG.1A) may be aligned within the first member 204 (e.g., via the first retaining mechanism 208) prior to closing the first member 204 such that a desired region of interest of the sample is aligned with the 43 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC barcoded array of the second substrate (e.g., the slide 104 from FIG.1A), e.g., when the first and second substrates are aligned in the sandwich configuration.
  • Such alignment may be accomplished manually (e.g., by a user) or automatically (e.g., via an automated alignment mechanism).
  • spacers may be applied to the first substrate 206 and/or the second substrate 212 to maintain a minimum spacing between the first substrate 206 and the second substrate 212 during sandwiching.
  • the permeabilization solution e.g., permeabilization solution 305
  • the first member 204 may then close over the second member 210 and form the sandwich configuration.
  • Analytes or analyte derivatives e.g., intermediate agents; e.g., ligation products
  • the image capture device 220 may capture images of the overlap area between the biological sample and the capture probes on the array 106. If more than one first substrates 206 and/or second substrates 212 are present within the sample handling apparatus 200, the image capture device 220 may be configured to capture one or more images of one or more overlap areas. [000184] Provided herein are methods for delivering a fluid to a biological sample disposed on an area of a first substrate and an array disposed on a second substrate.
  • FIGs.3A-3C depict a side view and a top view of an exemplary angled closure workflow 300 for sandwiching a first substrate (e.g., slide 303) having a biological sample 302 and a second substrate (e.g., slide 304 having capture probes 306) in accordance with some exemplary implementations.
  • FIG.3A depicts the first substrate (e.g., the slide 303 including a biological sample 302) angled over (superior to) the second substrate (e.g., slide 304).
  • reagent medium e.g., permeabilization solution
  • FIG.3A depicts the reagent medium on the right hand side of side view, it should be understood that such depiction is not meant to be limiting as to the location of the reagent medium on the spacer.
  • FIG.3B shows that as the first substrate lowers, and/or as the second substrate rises, the dropped side of the first substrate (e.g., a side of the slide 303 angled toward the second substrate) may contact the reagent medium 305.
  • the dropped side of the first substrate may urge 44 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC the reagent medium 305 toward the opposite direction (e.g., towards an opposite side of the spacer 310, towards an opposite side of the first substrate relative to the dropped side).
  • the reagent medium 305 may be urged from right to left as the sandwich is formed.
  • the first substrate and/or the second substrate are further moved to achieve an approximately parallel arrangement of the first substrate and the second substrate.
  • FIG.3C depicts a full closure of the sandwich between the first substrate and the second substrate with the spacer 310 contacting both the first substrate and the second substrate and maintaining a separation distance and optionally the approximately parallel arrangement between the two substrates.
  • the spacer 310 fully encloses and surrounds the biological sample 302 and the capture probes 306, and the spacer 310 form the sides of chamber 350 which holds a volume of the reagent medium 305.
  • Another method is 47 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC to cleave spatially-barcoded capture probes from an array and promote the spatially-barcoded capture probes towards and/or into or onto the biological sample.
  • a second type of capture probe associated with the feature includes the spatial barcode 702 in combination with a random N-mer capture domain 704 for gDNA analysis.
  • a third type of capture probe associated with the feature includes the spatial barcode 702 in combination with a capture domain complementary to the analyte capture agent of interest 705.
  • a fourth type of capture probe associated with the feature includes the spatial barcode 702 in combination with a capture probe that can specifically bind a nucleic acid molecule 706 that can function in a 51 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC CRISPR assay (e.g., CRISPR/Cas9).
  • a perturbation agent can be a small molecule, an antibody, a drug, an aptamer, a miRNA, a physical environmental (e.g., temperature change), or any other known perturbation agents.
  • the functional sequences can generally be selected for compatibility with any of a variety of different sequencing systems, e.g., Ion Torrent Proton or PGM, Illumina ® sequencing instruments, PacBio, Oxford Nanopore, etc., and the requirements thereof. In some embodiments, functional sequences can be selected for compatibility with non-commercialized sequencing systems.
  • FIG.8 depicts an exemplary arrangement of barcoded features within an array.
  • FIG.8 shows (L) a slide including six spatially-barcoded arrays, (C) an enlarged schematic of one of the six spatially-barcoded arrays, showing a grid of barcoded features in 52 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC relation to a biological sample, and (R) an enlarged schematic of one section of an array, showing the specific identification of multiple features within the array (labelled as ID578, ID579, ID560, etc.).
  • one of the oligonucleotides includes at least two ribonucleic acid bases at the 3’ end and/or the other oligonucleotide includes a phosphorylated nucleotide at the 5’ end.
  • one of the two oligonucleotides includes a capture binding capture domain (e.g., a poly(A) sequence, a non-homopolymeric sequence).
  • a ligase e.g., a T4 RNA ligase (Rnl2), a PBCV-1 DNA Ligase or Chorella virus DNA Ligase, a single-stranded DNA ligase, or a T4 DNA ligase
  • a ligase e.g., a T4 RNA ligase (Rnl2), a PBCV-1 DNA Ligase or Chorella virus DNA Ligase, a single-stranded DNA ligase, or a T4 DNA ligase
  • the two oligonucleotides hybridize to 53 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC sequences that are not adjacent to one another.
  • hybridization of the two oligonucleotides creates a gap between the hybridized oligonucleotides.
  • a polymerase e.g., a DNA polymerase
  • the ligation product is released from the analyte.
  • the ligation product is released using an endonuclease (e.g., RNAse H).
  • the ligation product is removed using heat.
  • the ligation product is removed using KOH.
  • the released ligation product can then be captured by capture probes (e.g., instead of direct capture of an analyte) on an array, optionally amplified, and sequenced, thus determining the location and optionally the abundance of the analyte in the biological sample.
  • capture probes e.g., instead of direct capture of an analyte
  • FIG.9A A non-limiting example of templated ligation methods disclosed herein is depicted in FIG.9A.
  • a biological sample is contacted with a substrate including a plurality of capture probes and contacted with (a) a first probe 901 having a target-hybridization sequence 903 and a primer sequence 902 and (b) a second probe 904 having a target-hybridization sequence 905 and a capture domain (e.g., a poly-A sequence) 906, the first probe 901 and a second probe 904 hybridize 910 to an analyte 907.
  • a ligase 921 ligates 920 the first probe to the second probe thereby generating a ligation product 922.
  • the capture probe can also include a unique molecular identifier (UMI) 9007, a spatial barcode 9008, a functional sequence 9009, and a cleavage domain 9010.
  • UMI unique molecular identifier
  • methods provided herein include permeabilization of the biological sample such that the capture probe can more easily bind to the captured ligated probe (i.e., compared to no permeabilization).
  • RT reverse transcription
  • the cDNA can then be enzymatically fragmented and size-selected in order to optimize the cDNA amplicon size.
  • P59016, i59017, i79018, and P79019 and can be used as sample indexes, and TruSeqTM Read 2 can be added via End Repair, A-tailing, Adaptor Ligation, and PCR.
  • the cDNA fragments can then be sequenced using paired-end sequencing using TruSeqTM Read 1 and TruSeqTM Read 2 as sequencing primer sites.
  • detection of one or more analytes e.g., protein analytes
  • analyte capture agents e.g., protein analytes
  • an “analyte capture agent” refers to an agent that interacts with an analyte (e.g., an analyte in a biological sample) and with a capture probe (e.g., a capture probe attached to a substrate or a feature) to identify the analyte.
  • the analyte capture agent includes: (i) an analyte binding moiety (e.g., that binds to an analyte), for example, an antibody or antigen-binding fragment thereof; (ii) analyte binding moiety barcode; and (iii) an analyte capture sequence.
  • analyte binding moiety barcode refers to a barcode that is associated with or otherwise identifies the analyte binding moiety.
  • analyte capture sequence refers to a region or moiety configured to hybridize to, bind to, couple to, or otherwise interact with a capture domain of a capture probe.
  • an analyte binding moiety barcode (or portion thereof) may be able to be removed (e.g., cleaved) from the analyte capture agent. Additional description of analyte capture agents can be found in Section (II)(b)(ix) of PCT Publication No.
  • FIG.10 is a schematic diagram of an exemplary analyte capture agent 1002 comprised of an analyte-binding moiety 1004 and an analyte-binding moiety barcode domain 1008.
  • the exemplary analyte -binding moiety 1004 is a molecule capable of binding to an analyte 1006 and the analyte capture agent is capable of interacting with a spatially-barcoded capture probe.
  • the analyte -binding moiety can bind to the analyte 1006 with high affinity and/or with high specificity.
  • the analyte capture agent can include an analyte-binding moiety barcode domain 1008, a nucleotide sequence (e.g., an oligonucleotide), which can hybridize to at least a portion or an entirety of a capture domain of a capture probe.
  • the analyte-binding moiety barcode domain 1008 can comprise an analyte binding moiety barcode and a capture handle sequence described herein.
  • the analyte-binding moiety 1004 can include a polypeptide and/or an aptamer.
  • FIG.11 is a schematic diagram depicting an exemplary interaction between a feature- immobilized capture probe 1124 and an analyte capture agent 1126.
  • the feature-immobilized capture probe 1124 can include a spatial barcode 1108 as well as functional sequences 1106 and UMI 1110, as described elsewhere herein.
  • the capture probe can be affixed 1104 to a feature (e.g., bead) or array 1102.
  • the capture probe can also include a capture domain 1112 that is capable of binding to an analyte capture agent 1126.
  • the analyte capture agent 1126 can include a functional sequence 1118, analyte binding moiety barcode 1116, and a capture handle sequence 1114 that is capable of binding to the capture domain 1112 of the capture probe 1124.
  • the analyte capture agent can also include a linker 1120 that allows the capture agent barcode domain 1116 to couple to the analyte binding moiety 1122.
  • specific capture probes and the analytes they capture are associated with specific locations in an array of features on a substrate.
  • specific spatial barcodes can be associated with specific array 56 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC locations prior to array fabrication, and the sequences of the spatial barcodes can be stored (e.g., in a database) along with specific array location information, so that each spatial barcode uniquely maps to a particular array location.
  • specific spatial barcodes can be deposited at predetermined locations in an array of features during fabrication such that at each location, only one type of spatial barcode is present so that spatial barcodes are uniquely associated with a single feature of the array.
  • the arrays can be decoded using any of the methods described herein so that spatial barcodes are uniquely associated with array feature locations, and this mapping can be stored as described above.
  • sequence information is obtained for capture probes and/or analytes during analysis of spatial information, the locations of the capture probes and/or analytes can be determined by referring to the stored information that uniquely associates each spatial barcode with an array feature location.
  • Each array feature location represents a position relative to a coordinate reference point (e.g., an array location, a fiducial marker) for the array. Accordingly, each feature location has an “address” or location in the coordinate space of the array.
  • a coordinate reference point e.g., an array location, a fiducial marker
  • each feature location has an “address” or location in the coordinate space of the array.
  • spatial analysis can be performed using dedicated hardware and/or software, such as any of the systems described in Sections (II)(e)(ii) and/or (V) of PCT Publication No. WO2020/176788 and/or U.S.
  • Suitable systems for performing spatial analysis can include components such as a chamber (e.g., a flow cell or sealable, fluid-tight chamber) for containing a biological sample.
  • a chamber e.g., a flow cell or sealable, fluid-tight chamber
  • the biological sample can be mounted for example, in a biological sample holder.
  • One or more fluid chambers can be connected to the chamber and/or the sample holder via fluid conduits, and fluids can be delivered into the chamber and/or sample holder via fluidic pumps, vacuum sources, or other devices coupled to the fluid conduits that create a pressure gradient to drive fluid flow.
  • One or more valves can also be connected to fluid conduits to regulate the flow of reagents from reservoirs to the chamber and/or sample holder.
  • the systems can optionally include a control unit that includes one or more electronic processors, an input interface, an output interface (such as a display), and a storage unit (e.g., a solid state storage medium such as, but not limited to, a magnetic, optical, or other solid state, persistent, writeable and/or re-writeable storage medium).
  • the control unit can optionally be connected to one or more remote devices via a network.
  • the control unit (and components thereof) can generally perform any of the steps and functions described herein. Where the system is connected to a remote device, the remote device (or devices) can perform any of the steps or features described herein.
  • the systems can optionally include one or more detectors (e.g., CCD, CMOS) used to capture images.
  • the systems can also optionally include one or more light sources (e.g., LED-based, diode-based, lasers) for illuminating a sample, a substrate with features, analytes from a biological sample captured on a substrate, and various control and calibration media.
  • the systems can optionally include software instructions encoded and/or implemented in one or more of tangible storage media and hardware components such as application specific integrated circuits.
  • the software instructions when executed by a control unit (and in particular, an electronic processor) or an integrated circuit, can cause the control unit, integrated circuit, or other component executing the software instructions to perform any of the method steps or functions described herein.
  • the systems described herein can detect (e.g., register an image) the biological sample on the array.
  • Exemplary methods to detect the biological sample on an array 58 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC are described in PCT Publication No. WO2021/102003 and/or U.S. Patent Application Publication No.2021/0150707, each of which is incorporated herein by reference in their entireties.
  • the biological sample Prior to transferring analytes from the biological sample to the array of features on the substrate, the biological sample can be aligned with the array.
  • Alignment of a biological sample and an array of features including capture probes can facilitate spatial analysis, which can be used to detect differences in analyte presence and/or level within different positions in the biological sample, for example, to generate a three-dimensional map of the analyte presence and/or level.
  • Exemplary methods to generate a two- and/or three-dimensional map of the analyte presence and/or level are described in PCT Publication No. WO2020/053655 and spatial analysis methods are generally described in PCT Publication No. WO2021/102039 and/or U.S. Patent Application Publication No.2021/0155982, each of which is incorporated herein by reference in their entireties.
  • a map of analyte presence and/or level can be aligned to an image of a biological sample using one or more fiducial markers, e.g., objects placed in the field of view of an imaging system which appear in the image produced, as described in the Substrate Attributes Section, Control Slide for Imaging Section of PCT Publication Nos. WO2020/123320, WO 2021/102005, and/or U.S. Patent Application Publication No.2021/0158522, each of which is incorporated herein by reference in their entireties.
  • fiducial markers e.g., objects placed in the field of view of an imaging system which appear in the image produced, as described in the Substrate Attributes Section, Control Slide for Imaging Section of PCT Publication Nos. WO2020/123320, WO 2021/102005, and/or U.S. Patent Application Publication No.2021/0158522, each of which is incorporated herein by reference in their entireties.
  • Fiducial markers can be used as a point of reference or measurement scale for alignment (e.g., to align a sample and an array, to align two substrates, to determine a location of a sample or array on a substrate relative to a fiducial marker) and/or for quantitative measurements of sizes and/or distances.
  • Reverse transcriptases or reverse transcription (RT) enzymes are RNA-dependent DNA polymerases, typically used to create a copy of an RNA sequence thereby generating a cDNA molecule. Reverse transcription is initiated by hybridization of a priming sequence to an RNA molecule which is extended by a reverse transcription enzyme in a template directed fashion.
  • a reverse transcription enzyme adds a plurality of non-template nucleotides to a nucleotide strand, thereby producing complementary deoxyribonucleic acid (cDNA) molecules.
  • the resultant cDNA can then be dehybridized from the template RNA molecule in any number of ways as 59 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC known in the art.
  • Engineered and/or recombinant are used interchangeably with respect to reverse transcriptase (RT) variant and/or fusion RT.
  • a DNA binding domain is a protein, or a defined region of a protein, that binds to a nucleic acid in a sequence-independent matter. For example, binding of the protein to DNA does not exhibit any preference for a particular sequence.
  • the DNA binding domain may be single or double stranded.
  • the AT-rich interaction domain can be a 62 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC KRPR* repeat, or SEQ ID NO: 21.
  • the AT-rich interaction domain can be a K/RKRGRPKK repeat, or SEQ ID NO: 16.
  • the AT-rich interaction domain can be a mammalian high mobility group I protein (HMG-I, or a-protein) AT-rich interaction domain or SEQ ID NO: 15.
  • the AT-rich interaction domain can comprise a Drosophila melanogaster D1 protein AT-rich interaction domain-consensus domain or SEQ ID NO: 16.
  • the tag is a hexahistidine tag.
  • the tag is selected from a small ubiquitin-like modifier tag (SUMO), a VariFlex C-Terminal solubility enhancement tag, a short peptide C-terminal tag, Thioredoxin (Trx) tag, Solubility-enhancer peptide sequences (SET) tag, IgG domain B1 of Protein G (GB1) tag, IgG repeat domain ZZ of Protein A (ZZ) tag, Solubility enhancing Ubiquitous Tag (SNUT tag), Seventeen kilodalton protein (Skp tag), Phage T7 protein kinase (T7PK) tag, E.
  • SUMO small ubiquitin-like modifier tag
  • Trx VariFlex C-Terminal solubility enhancement tag
  • SET Solubility-enhancer peptide sequences
  • IgG domain B1 of Protein G GB1
  • ZZ Solubility enhancing Ubiquitous Tag
  • the affinity tag may include, but is not limited to, albumin binding protein (ABP), AU1 epitope, AU5 epitope, T7-tag, V5-tag, B-tag, Chloramphenicol Acetyl Transferase (CAT), Dihydrofolate reductase (DHFR), AviTag, Calmodulin-tag, polyglutamate tag, E-tag, FLAG-tag, HA-tag, Myc-tag, NE-tag, S-tag, SBP-tag, Doftag 1, Softag 3, Spot-tag, tetracysteine (TC) tag, Ty tag, VSV-tag, Xpress tag, biotin carboxyl carrier protein 70 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC (BCCP), green fluorescent protein tag, HaloTag, Nus-tag, thioredoxin-tag, Fc-tag, cellulose binding domain, chitin binding protein (ABP),
  • the protease cleavage sequence is recognized by a protease including, but not limited to, alanine carboxypeptidase, Armillaria mellea astacin, bacterial leucyl aminopeptidase, cancer procoagulant, cathepsin B, clostripain, cytosol alanyl aminopeptidase, elastase, endoproteinase Arg-C, enterokinase, gastricsin, gelatinase, Gly-X carboxypeptidase, glycyl endopeptidase, human rhinovirus 3C protease, hypodermin C, Iga-specific serine endopeptidase, leucyl aminopeptidase, leucyl endopeptidase, lysC, lysosomal pro-X carboxypeptidase, lysyl aminopeptidase, methionyl aminopeptidase, myxobacter
  • the protease cleavage sequence is a thrombin cleavage sequence.
  • D. Reverse Transcriptase Polypeptides Reverse transcriptases or reverse transcription enzymes are known in the art to perform a reverse transcription reaction. As used herein, “Reverse transcriptase” and “reverse transcription enzyme” are synonymous. Reverse transcription is initiated by hybridization of a priming sequence to an RNA molecule which is extended by an engineered reverse transcription 71 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC enzyme in a template directed fashion.
  • a reverse transcription enzyme adds a plurality of non- template oligonucleotides to a nucleotide strand.
  • the reverse transcription reaction can produce single stranded complementary deoxyribonucleic acid (cDNA) molecules each having a molecular tag on a 5’ end thereof, followed by amplification of cDNA to produce a double stranded DNA having the molecular tag on the 5’ end and a 3’ end of the double stranded DNA.
  • cDNA complementary deoxyribonucleic acid
  • wild-type refers to a gene or gene product that has the characteristics of that gene or gene product when isolated from a naturally occurring source.
  • the amino acid sequence set forth in SEQ ID NO: 7 is a wild-type MMLV amino acid sequence.
  • the RT polypeptide can comprise an amino acid sequence that is at least 90% identical to SEQ ID NO: 1.
  • the RT polypeptide can comprise an amino acid sequence that is 90-99.99% identical to SEQ ID NO: 1.
  • the RT polypeptide can comprise an amino acid sequence that is 92-99.99% identical to SEQ ID NO: 1.
  • the RT polypeptide can comprise an amino acid sequence that is 93-99.99% identical to SEQ ID NO: 1.
  • the RT polypeptide can comprise an amino acid sequence that is 94-99.99% identical to SEQ ID NO: 1. In some embodiments, the RT polypeptide can comprise an amino acid sequence that is 95-99.99% identical to SEQ ID NO: 1. In some embodiments, the RT polypeptide can comprise an amino acid sequence that is 96-99.99% identical to SEQ ID NO: 1. In some embodiments, the RT polypeptide can comprise an amino acid sequence that is 97-99.99% identical to SEQ ID NO: 1. In some embodiments, the RT polypeptide can comprise an amino acid sequence that is 98-99.99% identical to SEQ ID NO: 1.
  • the RT polypeptide can comprise an amino acid sequence that is 90%, 91%, 92%, 93%, 94%, 95%, 96%, 97%, 98%, 99% or 99.5% identical to SEQ ID NO: 1.
  • the amino acid variation are at any one position or combination thereof as identified in an alignment of SEQ ID NO: 1 to any one of the RT polypeptide sequences in Table 1, or Table 2.
  • the RT polypeptide can comprise the amino acid sequence set forth in SEQ ID NO: 7.
  • the engineered reverse transcriptase can exhibit an altered reverse 72 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC transcriptase activity as compared to a reverse transcriptase having the amino acid sequence set forth in SEQ ID NO: 1 or 7.
  • the RT polypeptide can be a variant MMLV reverse-transcriptase having one or more mutations.
  • the RT polypeptide contemplated by the present disclosure can comprise a combination of mutations in the amino acid sequence of either the wild-type MMLV (SEQ ID NO 7 or 178) or in a MMLV variant (SEQ ID NO: 1, 143 or 179).
  • the amino acid sequence of the RT polypeptide sequence contemplated by the present disclosure can be at least 90% identical to SEQ ID NO: 1 or 143.
  • the amino acid sequence of the RT polypeptide sequence can be about 90% to about 99.99% identical to SEQ ID NO: 1 or 143, about 92% to about 99.99% identical to SEQ ID NO: 1 or 143, about 93% to about 99.99% identical to SEQ ID NO: 1 or 143, about 94% to about 99.99% identical to SEQ ID NO: 1 or 143, about 95% to about 99.99% identical to SEQ ID NO: 1 or 143, about 96% to about 99.99% identical to SEQ ID NO: 1 or 143, about 97% to about 99.99% identical to SEQ ID NO: 1 or 143, or about 98% to about 99.99% identical to SEQ ID NO: 1 or 143.
  • the amino acid sequence of the RT polypeptide sequence can be about 90%, about 91%, about 92%, about 93%, about 94%, about 95%, about 96%, about 97%, about 98%, about 99% or about 99.5% identical to SEQ ID NO: 1 or 143.
  • the present disclosure relates to engineered RT polypeptides or recombinant RT polypeptide comprising a wild-type RT or modified reverse transcriptases that comprise one or more (e.g., one, two, three, four, five, ten, twelve, fifteen, twenty, etc.) amino 73 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC acid changes. These amino acid changes render the reverse transcriptase more efficient for nucleic acid synthesis (e.g., single cell profiling assay) requiring very small volume, as compared to an unmutated or an unmodified reverse transcriptase.
  • nucleic acid synthesis e.g., single cell profiling assay
  • amino acids identified may be deleted and/or replaced with one or a number of amino acid residues.
  • any one or more of the amino acids may be substituted with any one or more amino acid residues such as Ala, Arg, Asn, Asp, Cys, Gln, Glu, Gly, His, He, Leu, Lys, Met, Phe, Pro, Ser, Thr, Trp, Tyr, and/or Val.
  • the RT polypeptide described herein comprises the amino acid sequence of SEQ ID NO:7, and comprises a combination of mutations selected from E69K, L139P, E302R, T306K, W313F, T330P, or N454K; and one or more of M39V, P47L, M66L, F155Y, D200N, D200E, H204R, G429S, L435G, L435K, P448A, D449G, H503V, D524N, T542D, E545G, D583N, H594Q, L603W, L603F, E607K, E607G, P627S, H634Y, H638G, A644V, D653H, K658R or L671P.
  • the engineered polypeptide can comprise the amino acid sequence of SEQ ID NO:7, and comprises a combination of mutations selected from E69K, L139P, D200N, E302R, T306K, W313F, T330P, L435G, P448A, D449G, N454K, D524N, or L603W, and E607K and one or more of M39V, P47L, M66L, F155Y, H204R, G429S, H503V, T542D, E545G, D583N, H594Q, P627S, H634Y, H638G, A644V, D653H, K658R or L671P.
  • the RT polypeptide sequence can comprise an amino acid sequence that is at least 95% identical to SEQ ID NO:1, 7, or 179, and a combination of mutations indexed to SEQ ID NO:7 or 178 selected from a combination of variants consisting of a T542D mutation, a D583N mutation, an E607G mutation, an A644V mutation, a D653H mutation, and a K658R mutation.
  • the RT polypeptide sequence can comprise an amino acid sequence that is at least 95% identical to SEQ ID NO:1, 7, or 179, and a combination of mutations indexed to SEQ ID NO:7 or 178 selected from a combination of variants consisting of an E545G mutation, a D583N mutation, an H594Q mutation, an L603F mutation, and a S679P mutation.
  • the RT polypeptide comprises an amino acid sequence that is at least 90% identical to an amino acid sequence selected from: SEQ ID NO: 14, SEQ ID NO: 22, SEQ ID NO: 23, SEQ ID NO: 24, SEQ ID NO: 25, SEQ ID NO: 26, SEQ ID NO: 27, SEQ ID NO: 28, SEQ ID NO: 29, SEQ ID NO: 30, SEQ ID NO: 31, SEQ ID NO: 32, SEQ ID NO: 33, SEQ ID NO: 35, SEQ ID NO: 36, SEQ ID NO: 37, SEQ ID NO: 38, SEQ ID NO: 39, SEQ 74 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC ID NO: 40, SEQ ID NO: 41, SEQ ID NO: 42, SEQ ID NO: 43, SEQ ID NO: 44, SEQ ID NO: 45, SEQ ID NO: 46, SEQ ID NO: 47, SEQ ID NO: 48, SEQ ID NO: 49,
  • the RT polypeptide can comprise SEQ ID NO: 63, 65, 67, 69, 71, 73, 75, 77, 79, 81, 83, 85, 87, 89, 91, 93, 95, 97, 99, 101, 103, 105, 107, 109, 111, 113, 115, 117, 119, 121, 123, 125, 127, 129, 131, 133, 141, 143, 145, 147, 149, 151, 157, 159, 172, or 173.
  • the RT polypeptide can comprise an amino acid sequence listed in Table 1 or 2.
  • the amino acid sequence of the RT polypeptide can also comprise E69K, L139P, D200N, E302R, T306K, W313F, T330P, N454K, H503V, D524N, L603W, E607K, and H634Y.
  • the amino acid sequence of the RT polypeptide comprises a combination of mutations selected from: M66L and L435G; M39V, M66L, and L435K; M39V and L435K; M66L, L435G, P448A and D449G; M39V, M66L, L435G, P448A and D449G; or M66L.
  • the amino acid sequence of the RT polypeptide comprises E69K, L139P, D200N, E302R, T306K, W313F, T330P, L435G, P448A, D449G, N454K, D524N, L603W, and E607K; and further comprises a combination of mutations selected from M66L; M66L and H503V; M66L and H634Y; and M66L, H503V, or H634Y.
  • the amino acid sequence of the RT polypeptide comprises M39V, E69K, L139P, a D200 mutation, E302R, T306K, W313F, T330P, G429S, P448A, a D449 mutation, L435K, N454K, a L603 mutation, a E607 mutation, and L671P and further comprises a second combination of mutations selected from D524N, T542D, P627S, A644V, D653H, or K658R mutation.
  • the D200 mutation is a D200N mutation
  • the D449 mutation is a D449G
  • the L603 mutation is an L603W
  • the E607 mutation is an E607G mutation.
  • the amino acid sequence of the RT polypeptide comprises M39V, E69K, L139P, a D200 mutation, E302R, T306K, W313F, T330P, G429S, P448A, a D449 mutation, L435K, N454K, a L603 mutation, a E607 mutation, and L671P and further comprises D524N, T542D, A644V, D653H, an R650H and K658R.
  • the D200 mutation is a D200N mutation
  • the D449 mutation is a D449G mutation
  • the L603 mutation is an L603F mutation
  • the E607 mutation is an E607K mutation.
  • the amino acid sequence of the RT polypeptide comprises M39V, E69K, L139P, a D200 mutation, E302R, T306K, W313F, T330P, G429S, P448A, a D449 mutation, L435K, N454K, a L603 mutation, a E607 mutation, and L671P and further comprises D524N, T542D, A644V, D653H, and K658R.
  • the D200 mutation is a D200N mutation
  • the D449 mutation is a D449E mutation
  • the L603 mutation is an L603W mutation
  • the E607 mutation is an E607G mutation.
  • the amino acid sequence of the RT polypeptide comprises M39V, E69K, L139P, a D200 mutation, E302R, T306K, W313F, T330P, G429S, P448A, a D449 mutation, L435K, N454K, a L603 mutation, a E607 mutation, and L671P and further comprises H204R, D524N, T542D, P627S, D583N, A644V, D653H and K658R.
  • the D200 mutation is a D200E mutation
  • the D449 mutation is a D449G mutation
  • the L603 mutation is an L603W mutation
  • the E607 mutation is an E607G mutation.
  • the amino acid sequence of the RT polypeptide comprises M39V, E69K, L139P, a D200 mutation, E302R, T306K, W313F, T330P, G429S, P448A, a D449 mutation, L435K, N454K, a L603 mutation, a E607 mutation, and L671P and further comprises H204R, E545G, D583N, and H594Q.
  • the D200 mutation is a D200E mutation
  • the D449 mutation is a D449G mutation
  • the L603 mutation is an L603F mutation
  • the E607 mutation is an E607K mutation.
  • the amino acid sequence of the RT polypeptide comprises M39V, E69K, L139P, a D200 mutation, E302R, T306K, W313F, T330P, G429S, P448A, a D449 mutation, L435K, N454K, a L603 mutation, a E607 mutation, and L671P and further comprises P47L, D524N, T542D, D583N, P627S, A644V, D653H, and K658R.
  • the D200 mutation is a D200N mutation
  • the D449 mutation is a D449G mutation
  • the L603 mutation is an L603W mutation
  • the E607 mutation is an E607G mutation.
  • the RT polypeptide comprises an amino acid sequence that is at least 95% identical to SEQ ID NO: 1 and the amino acid sequence of the engineered reverse transcriptase comprises at least one mutation indexed to SEQ ID NO:7 selected from a M17 mutation; an A32 mutation, a M44 mutation, a M39 mutation, a K47 mutation, a P51 mutation, an M66 mutation, an S67 mutation, an E69 mutation, a L72 mutation, a W94 mutation, a K103 mutation, an R110 mutation, a P117 mutation, an L139 mutation, an F155 mutation, an N178 mutation, an E179 mutation
  • the RT polypeptide comprises an amino acid sequence that is at least 95% identical to SEQ ID NO: 1 and the amino acid sequence of the engineered reverse transcriptase comprises an M39 mutation, a K47 mutation, an L435 mutation, a D449 mutation, a D524 mutation, an E607 mutation, a D653 mutation, and an L671 mutation as indexed to SEQ ID NO:7 and comprising at least one mutation indexed to SEQ ID NO:7 selected from a M17 mutation; an A32 mutation, a M44 mutation, a M39V mutation, a P51 mutation, an M66 mutation, an S67 mutation, an E69 mutation, a L72 mutation, a W94 mutation, a K103 mutation, an R110 mutation, a P117 mutation, an L139 mutation, an F155 mutation, an N178 mutation, an E179 mutation, a T197 mutation, a D200 mutation, an E201 mutation, an H204 mutation, a Q221 mutation, a V2
  • the engineered RT polypeptide exhibits an altered reverse transcriptase related activity when compared to a reverse transcriptase having the amino acid sequence set forth in SEQ ID NO: 1.
  • an RT polypeptide comprises an amino acid sequence that is at least 95% identical to SEQ ID NO: 1.
  • the engineered reverse transcriptase exhibits an altered reverse transcriptase related activity as compared to a reverse transcriptase having the amino acid sequence set forth in SEQ ID NO: 1.
  • the RT polypeptide comprises a combination of mutations indexed to SEQ ID NO:7 selected from: (i) an E69K mutation, an E302R mutation, a T306K mutation, a W313F mutation, a L435G mutation, or an N454K mutation, and comprising at least one mutation selected from an M39V mutation, an M66L mutation, an L139P mutation, an F155Y mutation, a D200N mutation, an E201Q mutation, a T287A mutation, a T330P mutation, an R411F mutation, a P448A mutation, a D449G mutation, an H503V mutation, an H594K mutation, L603W mutation, an E607K mutation, an H634Y mutation, a G637R mutation and an H638G mutation; (ii) an L139P mutation, a D200N mutation, a T330P mutation, an L603W mutation, or an E607K mutation, and comprising at least
  • the RT polypeptide comprises an amino acid sequence that is at least 95% identical to a reverse transcriptase having the amino acid sequence set forth in SEQ ID NO:1.
  • the engineered reverse transcription enzyme comprises an amino acid sequence that is at least 95% identical to SEQ ID NO: 1 and has at least one mutation selected from the group consisting of an M39V mutation, a P47L mutation, M66L mutation, an E69K mutation, an L139P mutation, a D200N mutation, an H204R mutation, an E302R mutation, a T306K mutation, a W313F mutation, a T330P mutation, an L435G mutation, a G429S mutation, an L435K mutation, a P448A mutation, a D449G mutation, a N454K mutation, an H503V mutation, a D524N mutation, a T542 mutation, an E545G mutation, a D583N mutation, an H594Q mutation
  • the disclosure provides an engineered RT polypeptide comprising an amino acid sequence that is at least 95% identical to SEQ ID NO:1 and the amino acid sequence of the engineered reverse transcriptase comprises a combination of mutations indexed to SEQ ID NO:7 or 178 selected from the group consisting of (a) an E69K mutation, an L139P mutation, a D200N mutation, an E302R mutation, a T306K mutation, a W313F mutation, a T330P mutation, a N454K mutation, an H503V mutation, a D524N mutation, an L603W mutation, an E607K mutation, and an H634Y mutation; (b) an M66L mutation, an E69K mutation, an L139P mutation, a D200N mutation, an E302R mutation, a T306K mutation, a W313F mutation, a T330P mutation, a N454K mutation, a D524N mutation, an H50
  • the D200 mutation is selected from the group consisting of D200N and D200E.
  • the D449 mutation is selected from the group consisting of D449G an D449E.
  • the L603 mutation is selected from the group consisting of L603W and L603F.
  • the E607 mutation is selected from the group consisting of E607G and E607K.
  • the engineered RT polypeptide further comprises at least one mutation selected from the group consisting of P47L, H204R, D524N, T542D, E545G, D583N, H594Q, P627S, A644V, R650H, D653H, K658R, L671P, and S679P.
  • an engineered reverse transcriptase of the present application has an amino acid sequence that is at least 95% identical to SEQ ID NO:1 and the amino acid sequence of said engineered reverse transcriptase comprises a combination of mutations indexed to SEQ ID NO:7 or 178; and the amino acid sequence of said engineered reverse transcriptase comprises a combination of mutations selected from the group consisting of: an E69K mutation, an L139P mutation, a D200N mutation, an E302R mutation, a T306K mutation, a W313F mutation, a T330P mutation, a N454K mutation, an H503V mutation, a D524N mutation, an L603W mutation, an E607K mutation, and an H634Y mutation and further comprising a second combination of mutations selected from the group consisting of: (a) an M66L mutation and an L534G mutation, (b) an M39V mutation, an M66L mutation and an L435K mutation, (
  • an engineered reverse transcriptase of the present application has an amino acid sequence that is at least 95% identical to SEQ ID NO:1 and the amino acid sequence of the engineered reverse transcriptase comprises a combination of mutations selected from the group consisting of: an M39V mutation, an E69K mutation, an L139P mutation, a D200 mutation, an E302R mutation, a T306K mutation, a W313F mutation, a T330P mutation, a 80 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC G429S mutation a P448A mutation, a D449 mutation, an L435K mutation, a N454K mutation, an L603 mutation, an E607 mutation, and an L671P mutation and further comprising a second combination of mutations selected from the group consisting of: (a) a D524N mutation, a T542D mutation, an A644
  • the P47 mutation is a P47L mutation
  • the D200 mutation is a D200N mutation
  • the D449 mutation is a D449G mutation
  • the L603 mutation is an L603W mutation
  • the E607 mutation is an E607G mutation
  • the P627 mutation is a P627S mutation.
  • a variant may comprise a first combination of mutations or alterations and may comprise an additional or second combination of mutations.
  • a first combination of mutations or alterations may include, but is not limited to, a combination of: (1) a M39 mutation, a K47 mutation, an L435 mutation, a D449 mutation, a D524 mutation, an E607 mutation, a D653 mutation and an L671 mutation; (2) an M39V mutation, a K47 mutation, an L435K mutation, a D449G mutation, a D524N mutation, an E607 mutation, a D653 mutation and an L671 mutation; (3) an M39 mutation, an M66 mutation, an E302 mutation, a T306 mutation, an L435 mutation, a D449 mutation, a D524 mutation, an E607 mutation, a D653 mutation and an L671 mutation; (4) an M39 mutation, an M66 mutation, an E302 (K or R) mutation
  • the second combination of mutations in a first engineered reverse transcriptase may comprise either a different set of mutations or a partially different second set of mutations as in a second engineered reverse transcriptase.
  • a second combination of mutations or alterations may include but is not limited to: (a) one or more mutations selected from an M17 mutation; an A32 mutation, a M44 mutation, a P51 mutation, an M66 mutation, an S67 mutation, an E69 mutation, a L72 mutation, a W94 mutation, a K103 mutation, an R110 mutation, a P117 mutation, an L139 mutation, an F155 mutation, an N178 mutation, an E179 mutation, a T197 mutation, a D200 mutation, an E201 mutation, an H204 mutation, a Q221 mutation, a V223 mutation, a V238 mutation, a G248 mutation, a T265 mutation, an E268 mutation, an R279 mutation, an R280 mutation, a K284 mutation, a T287 mutation,
  • the second combination of mutations may comprise a group of mutations as described herein and one or more additional mutations.
  • the engineered RT variants of the present disclosure comprise a M39V, M66I, Q91R, I347V, H594Q, or a combination thereof in the RT backbone of SEQ ID NO: 143 or SEQ ID NO: 7.
  • the engineered RT polypeptide comprises: M39V, M66I, Q91R, I347V, and H594Q (SEQ ID NO: 129 , SOLD 034).
  • the engineered RT variants of the present disclosure comprise M39V, T542D, D583N, E607G, A644V, D653H, K658R, L671P, or a combination thereof in the RT backbone of SEQ ID NO: 143 or SEQ ID NO: 7.
  • the engineered RT polypeptide comprises: M39V, T542D, D583N, E607G, A644V, D653H, K658R, and L671P (SEQ ID NO: 111, SOLD 025).
  • the engineered RT polypeptide comprises: M39V, M66I, Q91R, I347V, H594Q, or a combination thereof, and optionally M39V, M66I, Q91R, I347V, H594Q, or the combination thereof (substituted) in the RT sequence of SEQ ID NO: 143 (SEQ ID NO: 129, SOLD 034).
  • the engineered RT polypeptide comprises: M39V, T542D, D583N, E607G, A644V, D653H, K658R, L671P, or a combination thereof, and optionally M39V, T542D, D583N, E607G, A644V, D653H, K658R, L671P, or the combination thereof substituted in the RT sequence of SEQ ID NO: 143 (or SEQ ID NO: 7) (SEQ ID NO: 111, SOLD 025).
  • the engineered RT polypeptide comprises SOLD 33 VDG or SEQ ID NO: 173.
  • the engineered RT polypeptide described herein comprises an amino acid sequence that is at least about 90% identical to an amino acid sequence selected from the group consisting of SEQ ID NO: 27-61, 63, 65, 67, 69, 71, 73, 75, 77, 79, 81, 83, 85, 87, 89, 91, 93, 95, 97, 99, 101, 103, 105, 107, 109, 111, 113, 115, 117, 119, 121, 123, 125, 127, 129, 131, 133, 141, 143, 145, 147, 149, 151, 157, 159, 172, and 173.
  • the RT polypeptide is 42B L (SEQ ID NO: 145), 50A+G (SEQ ID NO: 147), SOLD 022 (SEQ ID NO: 105), SOLD 023 (SEQ ID NO: 107), SOLD 025 (SEQ ID NO: 111), SOLD 031 (SEQ ID NO: 123), SOLD 033 (SEQ ID NO: 127), SOLD 034 (SEQ ID NO: 129), SOLD 035 (SEQ ID NO: 131), SOLD 001 (SEQ ID NO: 65), SOLD 33 VDG (SEQ ID NO: 173), or an RT polypeptide set forth in SEQ ID NO: 143, or SEQ ID NO: 172.
  • E. Engineered DAT1 Reverse Transcriptases [000309]
  • One aspect of the present disclosure provides an engineered RT polypeptide comprising an amino acid sequence that is at least 90%, at least 92%, at least 95%, at least 97%, at least 98%, or at least 99% identical to an amino acid sequence of an RT disclosed in Table 1, or Table 2, and a DNA binding domain comprising an amino acid sequence selected from the group consisting of SEQ ID NO: 2, 3, 5, 6, 8, and 9.
  • Another aspect of the present disclosure provides an engineered RT polypeptide comprising an amino acid sequence that is at least 90%, at least 92%, at least 95%, at least 97%, at least 98%, or at least 99% identical to an amino acid sequence of SEQ ID NOs: 27-61, 63, 65, 67, 69, 71, 73, 75, 77, 79, 81, 83, 85, 87, 89, 91, 93, 95, 97, 99, 101, 103, 105, 107, 109, 111, 113, 115, 117, 119, 121, 123, 125, 127, 129, 131, 133, 141, 143, 145, 147, 149, 151, 157, 159, 172, or 173; and a DNA binding domain comprising an amino acid sequence selected from the group consisting of SEQ ID NO: 2, 3, 5, 6, 8, and 9.
  • Another aspect of the present disclosure provides an engineered RT polypeptide comprising an amino acid sequence that is at least 90%, at least 92%, at least 95%, at least 97%, at least 98%, or at least 99% identical to an amino acid sequence of SEQ ID NOs: 27-61, 63, 65, 67, 69, 71, 73, 75, 77, 79, 81, 83, 85, 87, 89, 91, 93, 95, 97, 99, 101, 103, 105, 107, 109, 111, 113, 115, 117, 119, 121, 123, 125, 127, 129, 131, 133, 141, 143, 145, 147, 149, 151, 157, 159, 172, or 173; and a DNA binding domain comprising an amino acid sequence selected
  • an engineered RT polypeptide comprising an amino acid sequence of an RT disclosed in Table 1 or Table 2; and an amino acid sequence of DNA binding domain disclosed in Table 1.
  • the engineered RT polypeptide described herein comprises the amino acid sequence of any one of SEQ ID NO: 174-188.
  • the engineered RT polypeptide described herein can comprise 42B L RT (SEQ ID NO: 145) operably linked to a full length DAT1 at the N-terminus.
  • the engineered RT can comprise the amino acid sequence of SEQ ID NO: 183.
  • the engineered RT polypeptide described herein can comprise 42B L RT operably linked to a full length DAT1 at the C-terminus.
  • the engineered RT can comprise the amino acid sequence of SEQ ID NO: 184.
  • DAT1(90) can be fused to any reverse transcriptase enzymes described herein, such as variants of MMLV reverse transcriptases disclosed in Table 1 or Table 2. While the majority of engineered RT polypeptide embodiments disclosed herein are N-terminus fusion proteins, the present disclosure also contemplates engineered RT polypeptides where the DNA binding domain, e.g., DAT1 or variant thereof is operably linked to the C-terminus of any RT polypeptide described herein.
  • the present disclosure provides any combination of DNA binding domain and RT polypeptide described in Table 1 or Table 2.
  • Non-exhaustive list of possible engineered RT polypeptides or recombinant RT proteins can comprise reverse transcriptases fused to full length DAT protein, DAT1 truncated to 36 amino acids, or DAT1(90) homologs.
  • the present disclosure provides an engineered RT comprising the amino acid sequence of SEQ ID NO: 178 (e.g., N-DAT-9042B).
  • the engineered RT can comprise the amino acid sequence of SEQ ID NO: 179 (e.g., N-DAT-9050A+ G).
  • the engineered RT can comprise the amino acid sequence of SEQ ID NO: 180 (e.g., N-DAT-90 SOLD 33 VDG).
  • the engineered RT can comprise the amino acid sequence of SEQ ID NO: 176 (e.g., N-DAT-90 SOLD 01).
  • the engineered RT can comprise the amino acid sequence of SEQ ID NO: 177 (e.g., C-DAT-90 SOLD 01).
  • an engineered RT polypeptide comprising N-DAT-9042B (SEQ ID NO: 178) and/or N-DAT-9050A+ G (SEQ ID NO: 178) were shown to enhance the sensitivity of the 5’ single cell assay as shown by enhanced transcript capture.
  • the engineered RT polypeptide described herein comprises 42B L RT (SEQ ID NO: 145) operably linked to an N-terminal truncated variant of DAT1 (D90) comprising the first 90 amino acids of the full length DAT1 at the C-terminus.
  • the engineered RT can comprise the amino acid sequence of SEQ ID NO: 174.
  • the variant can have at least 10%, 15%, 20%, 25%, 30%, 40%, 50%, 60%, 70%, 80%, 90%, 100% more or at least 2-fold, 3- fold, 4-fold, 5-fold, or 10-fold or more activity than the wild-type or known variant.
  • an engineered reverse transcription enzyme of the current application may exhibit an altered base-biased template switching activity such as an increased base-biased template switching activity, decreased base- biased template switching activity or an altered base-bias to the template switching activity.
  • the engineered RT polypeptide or the recombinant RT protein described herein exhibits increased template switching (TS) efficiency, increased processivity efficiency, increased binding affinity, increased transcription efficiency, increased chemical tolerance, improved ability to yield mitochondrial unique molecular identity (UMI) counts, improved ability to yield ribosomal unique molecular identity (UMI) counts, longer shelf life, higher strand displacement, higher end-to-end template jumping, or any combination thereof, when compared to an RT polypeptide or a recombinant RT protein lacking a conjugated DNA binding domain.
  • TS template switching
  • UMI mitochondrial unique molecular identity
  • UMI ribosomal unique molecular identity
  • variants comprising a M39V or a M66L mutation that do not exhibit altered performance in the 5’ GEM assay may exhibit an altered processivity, an altered kd or both.
  • L435K mutants may improve thermostability in the presence of primer template.
  • L435K variants may exhibit a thermal denaturation profile similar to that of the wild-type protein.
  • L435K, P448 and D449 are residues in the connection domain; altering these residues may result in increased conformational flexibility. Additionally, the connection domain is thought to impact the conformational flexibility of the RNAse H domain. H503 and H634 occur within the RNAse H domain.
  • the H503V and H634Y variants may impact primer-template contacting, processivity or both primer-template contacting and processivity.
  • Some variants share the following alterations: (a) the combination of variants consisting of a T542D mutation, a D583N mutation, an E607G mutation, an A644V mutation, a D653H mutation, and a K658R mutation.
  • Some variants share the following alterations: (b) the combination of variants consisting of an E545G mutation, a D583N mutation, an H594Q mutation, an L603F mutation, and a S679P mutation.
  • These variants may further comprise additional alterations that may affect one or more reverse transcriptase related activities.
  • the combination of variants consisting of a T542D mutation, a D583N mutation, an E607G mutation, an A644V mutation, a D653H mutation, and a K658R mutation and the combination of variants consisting of an E545G mutation, a D583N mutation, an H594Q mutation, an L603F mutation and a S679P mutation may exhibit an altered RNAse H activity.
  • RNase H activity [000340]
  • the engineered reverse transcriptase polypeptide or recombinant RT protein is engineered to have reduced and/or abolished RNase activity.
  • engineered reverse transcriptases or recombinant RT proteins of the disclosure are preferably modified or mutated such that the transcription efficiency of the engineered RT polypeptide or recombinant RT protein is increased or enhanced.
  • engineered reverse transcription polypeptide or recombinant RT proteins of the present disclosure may have an unexpectedly greater ability to associate or bind to full-length transcripts (e.g., in T-cell receptor paired transcriptional profiling), as compared to that exhibited by an enzyme having the amino acid sequence set forth in SEQ ID NO:1 or non-DAT1 engineered RT.
  • An altered transcription efficiency may be an increased transcription efficiency or a decreased transcription efficiency as compared to the transcription efficiency of a reverse transcriptase having the amino acid sequence set forth in SEQ ID NO: 1.
  • Altered transcription efficiency may be at least .1X, 0.2X, 0.3X, 0.4X, 0.5X, 0.6X, 0.7X, 0.8X, 0.9X, 1X, 1.5X, 2X, 2.5X, 3X, 3.5X, 4X, 4.5X, 5X, 5.5X, 6X, 6.5X, 7X, 7.5X, 8X, 8.5X, 9X, 10X, 15X, 20X, 25X or at least 30X greater than the transcription efficiency of a reverse transcriptase having the amino acid sequence set forth in SEQ ID NO: 1, 143, 145, or 172 or non-DAT1 engineered RT.
  • Transcription efficiency may be calculated as the sum of the area under the curve for the elongation, elongation plus tail, incomplete template switching (TSO) and complete template switching (TSO) regions over the total area under the curve for all products. Transcription efficiency reflects all those products for which transcription was successfully completed. Template switching oligonucleotide efficiency may be calculated as the area under the curve for the complete template switching region over the total area under the curve for all full-length products.
  • An engineered reverse transcriptase may have an increased transcription efficiency, an increased TSO efficiency or both an increased transcription efficiency and an increased TSO efficiency. 4.
  • the engineered reverse transcriptase polypeptide or recombinant RT protein described herein possesses one or more of the following characteristics when compared to a wild-type polymerase and/or reverse transcriptase: increased thermostability; increased thermoreactivity; increased resistance to reverse transcriptase inhibitors; increased ability to reverse transcribe difficult templates; increased speed; increased processivity; increased specificity; enhanced polymerization activity; increased sensitivity, or any combination thereof.
  • Processivity is defined as the ability of a polymerase or reverse transcriptase to carry out continuous nucleic acid synthesis on a template nucleic acid without frequent dissociation.
  • DNA polymerase or reverse transcriptase alone produces short DNA product strand per binding event.
  • Most DNA polymerases or reverse transcriptases are intrinsically low-processivity enzymes. The low processivity of DNA polymerase or reverse transcriptase alone is insufficient for the timely replication of a large genome.
  • the polymerization activity of the engineered reverse transcriptase polypeptide or recombinant RT protein as described herein is enhanced by about 5%, about 10%, about 15%, about 20%, about 25%, about 30%, about 35%, about 40%, about 45%, about 50%, about 55%, about 60%, about 65%, about 70%, about 75%, about 80%, about 85%, about 90%, about 90%, or about 100% as compared to the wild-type reverse transcriptase.
  • the engineered reverse transcriptase enzyme or engineered reverse transcriptase polypeptide or recombinant RT protein reverse transcribes a RNA molecule having at least about 100, at least about 200, at least about 300, at least about 400, at least about 500, at least about 600, at least about 700, at least about 800, at least about 900, or at least about 1000 nucleotides.
  • the engineered reverse transcriptase polypeptide or recombinant RT protein reverse transcribes a RNA molecule that is at least about 1kb, at least about 2kb, at least about 3kb, at least about 4 kb, at least about 5 kb, at least about 6 kb, at least about 7 kb, at least about 8 kb, at least about 9 kb, at least about 10kb, at least about 11 kb, at least about 12 kb, at least about 13 kb, at least about 14kb, or at least about 15 kb.
  • the engineered reverse transcriptase polypeptide or recombinant RT protein reverse transcribes a RNA molecule that is at least about 7kb or at least about 8kb.
  • the increase in thermoreactivity, resistance to reverse transcriptase inhibitors, ability to reverse transcribe difficult templates, speed, processivity, specificity, or sensitivity of the engineered reverse transcriptase polypeptide or recombinant RT protein as described herein has is about 5%, about 10%, about 15%, about 20%, about 25%, about 30%, about 35%, about 40%, about 45%, about 50%, about 55%, about 60%, about 65%, about 70%, about 75%, about 80%, about 85%, about 90%, about 90%, or about 100% as compared to the wild-type polymerase.
  • the enhanced reverse transcriptase activity is an increased binding affinity and template switching efficiency as compared to a reverse transcriptase having the amino acid sequence set forth in SEQ ID NO:1, 143, 145, or 172 or non-DAT1 engineered RT.
  • the enhanced reverse transcriptase activity is an enhanced processivity as compared to the processivity of a reverse transcriptase having the amino acid sequence set forth in SEQ ID NO: 1, 143, 145, or 172 or non-DAT1 engineered RT.
  • Processivity relates to a reverse transcriptase’s ability to remain associated with the template while incorporating nucleotides.
  • Measurements of processivity may include but are not limited to the number of nucleotides incorporated in a single binding event of a reverse transcriptase molecule.
  • Processivity also relates to the affinity of the enzyme for the substrate; thus, an enzyme with increased processivity may be more resistant to the presence of an inhibitor.
  • NUCLEIC ACIDS AND EXPRESSION VECTORS A. Nucleic Acids [000356]
  • One aspect of the present disclosure provides an isolated nucleic acid molecule encoding the engineered reverse transcriptase, the recombinant RT protein or derivatives thereof as described herein.
  • One aspect of the present disclosure provides an isolated nucleic acid molecule encoding any of the engineered RT polypeptides or the recombinant RT protein described herein.
  • the engineered reverse transcriptase polypeptide or the recombinant RT protein disclosed herein can been coded by a nucleic acid set forth herein or readily derived in light of polypeptide information provided herein and known in the art.
  • the isolated nucleic acid molecule encoding the RT polypeptide can comprise a sequence selected from SEQ ID NO; 25; SEQ ID NO: 136, SEQ ID NO:138, SEQ ID NO:140, SEQ ID NO:142, SEQ ID NO:144, SEQ ID NO:146, SEQ ID NO:148, SEQ ID NO:150, SEQ ID NO:152, SEQ ID NO:154, SEQ ID NO:156, SEQ ID NO:158, SEQ ID NO:160, SEQ ID NO:162, SEQ ID NO:164, SEQ ID NO:167, SEQ ID NO:169, or SEQ ID NO: 171; or a non- limiting embodiment of a nucleic acid sequence of Table 1 or Table 2.
  • the reverse transcriptase polypeptides, or the DNA binding domains need not be encoded by any specific nucleic acid exemplified herein.
  • redundancy in the genetic code allows for variations in nucleotide codon sequences that nevertheless encode the same amino acid.
  • engineered polymerases of the present disclosure can be produced from nucleic acid sequences that are different from those set forth herein, for example, being codon optimized for a particular expression system. Codon optimization can be carried out, for example, as set forth in Athey et al., BMC Bioinformatics, 18:391-401 (2017).
  • Wild type nucleic acids may be isolated from naturally occurring sources to be used as starting material to generate novel polymerases.
  • nomenclature and the laboratory procedures in recombinant DNA technology described below are those well-known and commonly employed in the art. Standard techniques for cloning, DNA and RNA isolation, amplification and purification are known.
  • enzymatic reactions involving DNA ligase, DNA polymerase, restriction endonucleases are the 97 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC like are performed according to the manufacturer's specifications.
  • nucleic acids e.g., RT or DNA binding domain
  • the isolation of nucleic acids may be accomplished by a variety of techniques.
  • the nucleic acids of the present disclosure can be generated from the wild type sequences.
  • the wild type sequences are altered to create modified sequences.
  • Wild type molecules e.g., RT or DNA binding domain
  • Exemplary modification methods are site-directed mutagenesis, point mismatch repair, or oligonucleotide-directed mutagenesis.
  • a “vector” refers to a polynucleotide, which when independent of the host chromosome, is capable replication in a host organism.
  • Preferred vectors include plasmids and typically have an origin of replication.
  • Vectors can comprise, e.g., transcription and translation terminators, transcription and translation initiation sequences, and promoters useful for regulation of the expression of the particular nucleic acid.
  • the polymerases of the present disclosure can be expressed in a variety of host cells, including E. coli, other bacterial hosts, yeasts, filamentous fungi, and various higher eukaryotic cells such as the COS, CHO and HeLa cells lines and myeloma cell lines.
  • E. coli E. coli
  • yeasts yeasts
  • filamentous fungi various higher eukaryotic cells
  • COS COS
  • CHO and HeLa cells lines and myeloma cell lines eukaryotic cells
  • Techniques for gene expression in microorganisms are described in, for example, Smith, Gene Expression in Recombinant Microorganisms (Bioprocess Technology, Vol.22), Marcel Dekker, 1994.
  • bacteria examples include, but are not limited to, Escherichia, Enterobacter, Azotobacter, Erwinia, Bacillus, Pseudomonas, Klebsielia, Proteus, Salmonella, Serratia, Shigella, Rhizobia, Vitreoscilla, and Paracoccus.
  • Filamentous fungi that are useful as expression hosts include, for example, the following genera: Aspergillus, Trichoderma, Neurospora, Penicillium, Cephalosporium, Achlya, Podospora, Mucor, Cochliobolus, and Pyricularia.
  • yeast Synthesis of heterologous proteins in yeast is 98 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC well known and described in the literature.
  • C. Host cells Another aspect of the present disclosure provides a host cell transfected with the expression vector comprising the isolated nucleic acid encoding the engineered reverse transcriptase as described herein.
  • Eukaryotic expression systems for mammalian cells, yeast, and insect cells are well known in the art and are also commercially available.
  • vectors include Yeast Integrating plasmids (e.g., YIp5) and Yeast Replicating plasmids (the YRp series plasmids) and pGPD-2.
  • Expression vectors containing regulatory elements from eukaryotic viruses are typically used in eukaryotic expression vectors, e.g., SV40 vectors, papilloma virus vectors, and vectors derived from Epstein-Barr virus.
  • exemplary eukaryotic vectors include pMSG, pAV009/A+, pMTO10/A+, pMAMneo-5, baculovirus pDSVE, and any other vector allowing expression of proteins under the direction of the CMV promoter, SV40 early promoter, SV40 later promoter, metallothionein promoter, murine mammary tumor virus promoter, Rous sarcoma virus promoter, polyhedrin promoter, or other promoters shown effective for expression in eukaryotic cells.
  • the engineered reverse transcriptase or a derivative thereof can be purified according to standard procedures of the art, including ammonium sulfate precipitation, affinity purification columns, column chromatography, gel electrophoresis and the like. Substantially pure compositions of at least about 90 to about 95% homogeneity are preferred, and about 98 to about 99% or more homogeneity are most preferred. Once purified, partially or to homogeneity as desired, the polypeptides may then be used (e.g., as immunogens for antibody production).
  • the nucleic acids that encode the engineered reverse transcriptase or derivatives thereof can also include a coding sequence for an epitope or “tag” for which an affinity binding reagent is available.
  • suitable epitopes include the myc and V-5 reporter genes; expression vectors useful for recombinant production of fusion polypeptides having these epitopes are commercially available (e.g., Invitrogen (Carlsbad Calif.) vectors pcDNA3.1/Myc-His and 99 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC pcDNA3.1/V5-His are suitable for expression in mammalian cells).
  • Additional expression vectors suitable for attaching a tag to the fusion proteins of the disclosure, and corresponding detection systems are known to those of skill in the art as described herein, and several are commercially available (e.g., FLAG′′ (Kodak, Rochester N.Y.).
  • FLAG′′ Kodak, Rochester N.Y.
  • Another example of a suitable tag is a polyhistidine sequence, which is capable of binding to metal chelate affinity ligands. Typically, six adjacent histidines are used (6His-tag, his-tag), although one can use more or less than six.
  • Suitable metal chelate affinity ligands that can serve as the binding moiety for a polyhistidine tag include nitrilo-tri-acetic acid (NTA).
  • the engineered reverse transcriptase or derivatives thereof may possess a conformation substantially different than the native conformations of the constituent polypeptides. In this case, it may be necessary or desirable to denature and reduce the engineered reverse transcriptase or a derivative thereof and cause the engineered reverse transcriptase or a derivative thereof to re-fold into the preferred conformation. Methods of reducing and denaturing proteins and inducing re- folding are well known to those of skill in the art. V.
  • compositions comprising a variety of components in various combinations needed for nucleic acid amplification using the engineered RT polypeptides or recombinant proteins disclosed herein.
  • One aspect of the present disclosure provides a composition comprising any of the recombinant RT proteins described herein.
  • One aspect of the present disclosure provides a composition comprising any of the engineered RT polypeptides described herein.
  • One aspect of the present disclosure provides a composition comprising any of the engineered RT polypeptides described herein.
  • composition comprising any of the expression vectors described herein.
  • any one of the compositions described herein further comprise a buffer.
  • the compositions are formulated by admixing one or more engineered reverse transcriptase polypeptides or recombinant RT proteins, 100 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC or derivatives thereof of the present disclosure in a buffered salt solution.
  • One or more DNA polymerases and/or one or more nucleotides, and/or one or more primers may optionally be added to create the compositions of the invention.
  • compositions can be used in the methods disclosed herein to produce, analyze, quantitate and otherwise manipulate nucleic acid molecules (e.g., using reverse transcription or one-step RT-PCR procedures).
  • the engineered reverse transcriptases or the recombinant RT proteins disclosed herein are provided at working concentrations (e.g., 1 ⁇ ) in stable buffered salt solutions.
  • working concentration means the concentration of an enzyme (e.g., engineered reverse transcriptase or the recombinant RT protein) that is at or near the optimal concentration used in a solution to perform a particular function such as reverse transcription of nucleic acids.
  • Such compositions can also be formulated as concentrated stock solutions (e.g., 2 ⁇ , 3 ⁇ , 4 ⁇ , 5 ⁇ , 6 ⁇ , 10 ⁇ , etc.). In some embodiments, having the composition as a concentrated (e.g., 5x) stock solution allows a greater amount of nucleic acid sample to be added (such as, for example, when the compositions are used for nucleic acid synthesis).
  • the water used in forming the compositions of the present invention is preferably distilled, deionized and sterile filtered (through a 0.1-0.2 micrometer filter) and is free of contamination by DNase and RNase enzymes.
  • Such water is available commercially, for example from Life Technologies (Carlsbad, Calif.) or may be made as needed according to methods well known to those skilled in the art.
  • METHODS FOR USING ENGINEERED REVERSE TRANSCRIPTASES [000373]
  • the engineered reverse transcriptases of the present disclosure may be used in any application in which a reverse transcriptase with the indicated altered activity is desired.
  • One aspect of the present disclosure provides a method for performing a reverse transcription reaction for generating a nucleic acid product from an RNA template using an engineered reverse transcriptase or recombinant RT protein described herein.
  • the engineered reverse transcriptases or recombinant RT protein of the present application may be used in any application in which a reverse transcriptase with the indicated altered activity is desired. Methods of using reverse transcriptases are known in the art. One skilled in the art may select any of the engineered reverse transcriptases disclosed herein.
  • One aspect of the present disclosure provides a method for performing a reverse transcription reaction for generating a nucleic acid product from an RNA template comprising contacting under suitable conditions a biological sample or extract thereof with an engineered RT polypeptide, or a recombinant RT protein described herein.
  • the cell can be fixed.
  • the cell can be permeabilized.
  • the cell can be permeabilized and fixed.
  • the cell is a cell bead.
  • the cell bead is fixed.
  • the nucleus is permeabilized or fixed.
  • the nucleus is permeabilized and fixed.
  • the biological sample comprises a suitable cellular preparation selected from cell populations and/or single cells.
  • the biological sample comprises a tissue. 102 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC [000376]
  • the reverse transcription reaction is part of a single cell RNA sequencing assay.
  • the single cell RNA sequencing assay further comprises, prior to the reverse transcription, partitioning the cell, the cell bead, or the nucleus into a partition.
  • the single cell RNA sequencing assay further comprises, after the reverse transcription reaction, hybridizing the nucleic acid product to an oligonucleotide molecule comprising a partition-specific barcode.
  • the reverse transcription reaction is part of a spatial RNA sequencing assay.
  • the engineered RT polypeptide or the recombinant RT protein enhances template switching (TS) efficiency, processivity efficiency, binding affinity, transcription efficiency, chemical tolerance, ability to yield mitochondrial unique molecular identity (UMI) counts, ability to yield ribosomal unique molecular identity (UMI) counts, strand displacement, end-to-end template jumping, or any combination thereof, when compared to an RT polypeptide or a recombinant RT protein lacking a conjugated DNA binding domain.
  • TS template switching
  • UMI mitochondrial unique molecular identity
  • UMI ribosomal unique molecular identity
  • the engineered RT polypeptide or the recombinant RT protein enhances at least two or more of template switching (TS) efficiency, processivity efficiency, binding affinity, transcription efficiency, chemical tolerance, ability to yield mitochondrial unique molecular identity (UMI) counts, strand displacement, end-to-end template jumping, or ability to yield ribosomal unique molecular identity (UMI) counts, when compared to an RT polypeptide or a recombinant RT protein lacking a conjugated DNA binding domain.
  • TS template switching
  • UMI mitochondrial unique molecular identity
  • UMI ribosomal unique molecular identity
  • the engineered RT polypeptide or the recombinant RT protein can comprise a DNA binding domain comprising an amino acid sequence selected from SEQ ID NO:2, 3, 5, 6, 8, 9, or 11-24; and an amino acid sequence selected from SEQ ID NOs: 27-61, 63, 65, 67, 69, 71, 73, 75, 77, 79, 81, 83, 85, 87, 89, 91, 93, 95, 97, 99, 101, 103, 105, 107, 109, 111, 113, 115, 117, 119, 121, 123, 125, 127, 129, 131, 133, 141, 143, 145, 147, 149, 151, 157, 159, 172, or 173.
  • the amino acid sequence of the engineered RT polypeptide or the recombinant RT protein comprises an amino acid sequence having at least about 90%, at least about 92%, at least about 95%, at least about 97%, at least about 98%, or at least about 99% sequence identity to SEQ ID NO: 174-188.
  • the engineered RT polypeptide or the recombinant RT protein can comprise M39V, M66I, Q91R, I347V, and H594Q substitution in SEQ ID NO: 143.
  • the engineered RT polypeptide or the recombinant RT protein can comprise SEQ ID NO: 129 (SOLD 034).
  • the engineered RT polypeptide or the recombinant RT protein can comprise M39V, T542D, D583N, E607G, A644V, D653H, K658R, and/or L671P in SEQ ID NO: 143.
  • the engineered RT polypeptide or the recombinant RT protein can comprise SEQ ID NO: 111 (SOLD 025).
  • the engineered RT or the recombinant RT protein can comprise a M39V, M66I, Q91R, I347V, and/or H594Q in SEQ ID NO: 143.
  • the engineered RT polypeptide or the recombinant RT protein can comprise an amino acid sequence that is at least about 90%, at least about 92%, at least about 95%, at least about 97%, at least about 98%, or at least about 99% identical to an amino acid sequence disclosed in Table 1 or Table 2.
  • One aspect of the present disclosure provides a method of using the engineered RT polypeptide, or the recombinant RT protein described herein, the method comprising contacting the engineered RT polypeptide or the recombinant RT protein with a nucleic acid template under suitable conditions to produce a polymerized nucleic acid product.
  • the nucleic acid template comprises an RNA, a DNA, or a nucleic acid comprising an unnatural nucleotide.
  • the nucleic acid template comprises an RNA.
  • the engineered reverse transcriptase comprises an M39 mutation, a K47 mutation, an L435 mutation, a D449 mutation, a D524 mutation, an E607 mutation, a D653 mutation and an L671 mutation in SEQ ID NO:7.
  • the engineered reverse transcriptase comprises a mutation selected from a K13 mutation, a K13L mutation, a D36 mutation, an N37 mutation, a V2 mutation, a D36L mutation, an insertion, and a combination thereof.
  • the engineered reverse transcriptases or the recombinant RT proteins, or derivatives thereof of the present disclosure are used in reverse transcription reactions, such as RT-PCR, or other known reactions in the art where nucleic acids, for example RNA molecules, are reverse transcribed using a reverse transcriptase.
  • the engineered reverse transcriptase, the recombinant RT protein or a derivative thereof as described herein may be used to make nucleic acid molecules from one or more templates.
  • Such methods can comprise mixing one or more nucleic acid templates (e.g., RNA, such as non-coding RNA (ncRNA), messenger RNA (mRNA), micro RNA (miRNA), and small interfering RNA (siRNA) molecules) with one or more of the engineered reverse transcriptases of the disclosure and incubating the mixture under conditions sufficient to generate one or more nucleic acid molecules complementary to all or a portion of the one or more nucleic acid templates.
  • RNA such as non-coding RNA (ncRNA), messenger RNA (mRNA), micro RNA (miRNA), and small interfering RNA (siRNA) molecules
  • ncRNA non-coding RNA
  • mRNA messenger RNA
  • miRNA micro RNA
  • siRNA small interfering RNA
  • the method of using the engineered reverse transcriptase, or the recombinant RT protein or a derivative thereof as described herein comprises the amplification of one or more nucleic acid molecules comprising mixing one or more nucleic acid templates with one of the engineered reverse transcriptase polypeptide or recombinant RT 105 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC proteins or a derivative thereof of the disclosure, and incubating the mixture under conditions sufficient to amplify one or more nucleic acid molecules complementary to all or a portion of the one or more nucleic acid templates.
  • the method may comprise the use of one or more DNA polymerases and may be employed as in standard reverse transcription-polymerase chain reaction (RT-PCR) reactions.
  • RT-PCR reverse transcription-polymerase chain reaction
  • the method of using the engineered reverse transcriptase, recombinant RT protein or a derivative thereof as described herein may be one-step (e.g., one- step RT-PCR) or two-step (e.g., two-step RT-PCR) reactions.
  • the one-step RT-PCR type reactions may be accomplished in one tube thereby lowering the possibility of contamination.
  • Such one-step reactions can comprise (a) mixing a nucleic acid template (e.g., mRNA) with one or more engineered reverse transcriptase polypeptides or recombinant RT proteins or derivatives thereof of the present disclosure and one or more polymerases and (b) incubating the mixture under conditions sufficient to amplify a nucleic acid molecule complementary to all or a portion of the template.
  • a nucleic acid template e.g., mRNA
  • engineered reverse transcriptase polypeptides or recombinant RT proteins or derivatives thereof of the present disclosure e.g., RNA
  • Such methods can comprise (a) mixing a nucleic acid template (e.g., mRNA) with an engineered reverse transcriptase polypeptide or a recombinant RT protein or a derivative thereof of the present disclosure, (b) incubating the mixture under conditions sufficient to make a nucleic acid molecule (e.g., a DNA molecule) complementary to all or a portion of the template, (c) mixing the nucleic acid molecule with one or more DNA polymerases and (d) incubating the mixture of step (c) under conditions sufficient to amplify the nucleic acid molecule.
  • a nucleic acid template e.g., mRNA
  • an engineered reverse transcriptase polypeptide or a recombinant RT protein or a derivative thereof of the present disclosure incubating the mixture under conditions sufficient to make a nucleic acid molecule (e.g., a DNA molecule) complementary to all or a portion of the template
  • a nucleic acid molecule
  • a combination of DNA polymerases and the engineered reverse transcriptase polypeptide or recombinant RT protein or a derivative thereof of the present disclosure may be used.
  • Amplification methods which may be used in accordance with the present invention (e.g., using one or more engineered reverse transcriptase polypeptides or recombinant RT proteins or derivatives thereof of the present disclosure) include PCR, Isothermal Amplification, Strand Displacement Amplification (SDA), and Nucleic Acid Sequence-Based Amplification (NASBA); as well as more complex PCR-based nucleic acid fingerprinting techniques such as Random Amplified Polymorphic DNA (RAPD) analysis, Arbitrarily Primed PCR (AP-PCR) 106 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC DNA Amplification Fingerprinting (DAF); microsatellite PCR; Directed Amplification of Minisatellite-region DNA (DAVID); digital droplet PCT (ddPCR) and Amplification Fragment Length Polymorphism (AFLP) analysis.
  • RAPD Random Amplified Polymorphic DNA
  • the engineered reverse transcriptase disclosed herein may be used in methods of amplifying or sequencing a nucleic acid molecule comprising one or more polymerase chain reactions (PCRs), such as any of the PCR- based methods described above.
  • PCRs polymerase chain reactions
  • Methods of producing an engineered reverse transcriptase, an engineered reverse transcriptase or a derivative thereof of the present disclosure are known to those of skill in the art of molecular biology or molecular genetics.
  • nucleic acids encoding the wild-type polymerase or nucleic acid binding domains can be generated using routine techniques in the field of recombinant genetics.
  • nucleic Acid Sample Processing Another aspect of the present disclosure provides a nucleic acid extension method comprising contacting a target nucleic acid molecule with an engineered reverse transcriptase or a recombinant RT protein and a plurality of nucleic acid barcoded molecules comprising a barcode sequence, and incubating the target nucleic acid, the engineered reverse transcriptase or the recombinant RT protein and barcoded molecules under conditions in which the barcoded molecules are extended by the engineered reverse transcriptase or the recombinant RT protein.
  • the engineered reverse transcriptase or the recombinant RT protein comprises the amino acid sequence of an engineered RT or an recombinant RT protein described herein or a derivatives thereof.
  • the target nucleic acid hybridizes to one of the plurality of barcoded molecules and the hybridized barcoded molecule is extended by the engineered reverse transcriptase or the recombinant RT protein described herein.
  • the novel engineered reverse transcriptase polypeptide or the recombinant RT protein described herein can be used to generate a Single Cell 3' (SC-3') and/or 5’ (SC-5') gene expression libraries.
  • the SC-3' and SC-5' assays are similar but capture different ends of the polyadenylated transcript in the final library. Both solutions use poly-dT primer for reverse transcription).
  • the poly-dT sequence is located on the gel bead oligo.
  • the SC-5' assay (FIGs 14A) the poly-dT is supplied as an RT primer.
  • a template switching oligo (TSO) is used in both assays to reverse transcribe the full-length transcript.
  • transcripts are randomly fragmented under conditions that favor 300-400 bp length fragments. Downstream of fragmentation, only transcripts containing both (1) a 10x Barcode and (2) an Illumina ® Read 2 adaptor, which is ligated on to the cDNA after fragmentation, can be amplified during the Sample Index PCR.
  • the nucleic acid is a ribonucleic acid (RNA) molecule; and the engineered reverse transcriptase polypeptide or recombinant RT protein reverse transcribes the RNA molecule thereby generating a first strand cDNA.
  • a first strand cDNA reaction can be optionally performed using template switching oligonucleotides.
  • a template switching oligonucleotide can hybridize to a poly(C) tail added to a 3’ end of the cDNA by the engineered reverse transcriptase polypeptide or recombinant RT protein described herein.
  • the original mRNA template and template switching oligonucleotide can then be denatured from the cDNA and a barcoded capture probe can then hybridize with the cDNA and a complement of the cDNA can be generated.
  • the first strand cDNA can then be purified and collected for downstream amplification steps.
  • the first strand cDNA can be amplified using PCR, where the forward and reverse primers flank the spatial barcode and target analyte regions of interest, generating a library associated with a particular spatial barcode.
  • the cDNA comprises a sequencing by synthesis (SBS) primer sequence.
  • the library amplicons are sequenced and analyzed to decode spatial information.
  • a reverse transcription reaction introduces a barcode.
  • the barcoding reaction is an enzymatic reaction.
  • the barcoding reaction is a reverse transcription amplification reaction that generates complementary deoxyribonucleic acid (cDNA) molecules upon reverse transcription of ribonucleic acid (RNA) molecules of the cell.
  • RNA molecules are released from the cell.
  • the RNA molecules are released from the cell by lysing the cell.
  • the RNA molecules are released from the cell by permeabilizing the cell, or a tissue which comprises a plurality of the same and/or different cell types.
  • the RNA molecules are messenger RNA (mRNA).
  • a reverse transcription reaction using the engineered reverse transcriptase, the recombinant RT protein or derivative thereof of the present disclosure is initiated at the point of hybridization of the capture sequences to the RNA molecules, with the capture probe being extended by the engineered reverse transcriptase polypeptide or recombinant RT protein of the present disclosure in a template directed fashion using the hybridized mRNA as a template.
  • the recombinant RT protein or the engineered RT polypeptide can exhibit increased transcript capture during amplification.
  • the DNA binding domain of the engineered Rt polypeptide or recombinant RT protein can enhance the hybridization of a transcript and a primer during a nucleic acid amplification process.
  • the primer can comprise a poly-dT or a poly(dT)VN sequence and a non-poly(dT) sequence; and the transcript can comprise a poly-dA sequence.
  • the DNA binding domain e.g., DAT1 or variant thereof
  • the primer is a barcoded molecule.
  • the reverse transcription reaction produces single stranded cDNA molecules each having a molecular tag and barcode associated with the cDNA, followed by amplification of cDNA to produce a double stranded cDNA that includes the sequences of the barcoded molecules.
  • the plurality of nucleic acid barcoded molecules comprise an oligo(dT) sequence.
  • the engineered reverse transcriptase polypeptide or recombinant RT protein reverse transcribes the mRNA molecule into a complementary DNA molecule using the mRNA hybridized to the oligo(dT) sequence of the nucleic acid barcoded 109 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC molecules as a template, and the nucleic acid binding domain binds and stabilizes the mRNA- oligo(dT) hybrid during the reverse transcription.
  • the engineered reverse transcriptase polypeptide or recombinant RT protein as described herein further amplifies the complementary DNA molecule comprising the barcode sequence, thereby generating an amplified DNA product comprising the barcode sequence, molecular tag sequence, or complements thereof.
  • the method can comprise a second nucleic acid molecule comprising an oligo(dT) sequence.
  • the plurality of nucleic acid barcoded molecules comprise an oligo(dT) sequence; and the nucleic acid binding domain of the engineered reverse transcriptase polypeptide or recombinant RT protein binds and stabilizes the mRNA-Oligo(dT) hybrid, while the polymerase domain of the engineered reverse transcriptase polypeptide or recombinant RT protein reverse transcribes the mRNA molecule using the second nucleic acid molecule comprising the oligo(dT) sequence, thereby generating a complementary DNA molecule.
  • the engineered reverse transcriptase polypeptide or recombinant RT protein further amplifies the complementary DNA molecule, thereby generating an amplified DNA product comprising a barcode sequence.
  • the nucleic acid extension method comprises a cell, a population of cells, or a tissue and the template nucleic acid molecule is from the cell, population of cells or the tissue.
  • the molecular tags are coupled to priming sequences and the barcoding reaction is initiated by hybridization of the priming sequences to the RNA molecules.
  • each priming sequence comprises a random N-mer sequence.
  • the random N-mer sequence is complementary to a 3’ sequence of a ribonucleic acid molecule of the cell.
  • the random N-mer sequence comprises a poly- dT sequence having a length of at least 5 bases.
  • the random N-mer sequence comprises a poly-dT sequence having a length of at least 10 bases.
  • the barcoding reaction is performed by extending the priming sequences in a template directed fashion using reagents for reverse transcription.
  • the reagents for reverse transcription comprise a reverse transcription enzyme 110 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC (e.g., engineered RT polypeptide or recombinant RT protein), a buffer and a mixture of nucleotides.
  • the reverse transcription enzyme adds a plurality of non- template oligonucleotides upon reverse transcription of a ribonucleic acid molecule.
  • the reverse transcription enzyme is an engineered RT polypeptide or recombinant RT protein as disclosed herein.
  • the barcoding reaction produces single stranded complementary deoxyribonucleic acid (cDNA) molecules each having a molecular tag from said molecular tags on a 5’ end thereof, followed by amplification of cDNA to produce a double stranded cDNA having the molecular tag on the 5’ end and a 3’ end of the double stranded cDNA.
  • cDNA complementary deoxyribonucleic acid
  • a molecular tag which comprises a barcode plus additional functional sequences, or only additional functional sequences is further included into a cDNA molecule generated during a reverse transcription reaction.
  • the reagents for reverse transcription comprise a reverse transcription enzyme (e.g., the engineered reverse transcriptase or the recombinant RT protein described herein), a buffer, and a mixture of nucleotides.
  • the reverse transcription enzyme adds a plurality of non- template oligonucleotides upon reverse transcription of a ribonucleic acid molecule from the nucleic acid molecules.
  • the reverse transcription enzyme is an engineered reverse transcriptase or a recombinant RT protein as disclosed herein.
  • the present disclosure provides methods that utilize the engineered reverse transcriptase polypeptides or the recombinant RT protein described herein for nucleic acid sample processing.
  • the method comprises contacting a template ribonucleic acid (RNA) molecule with an engineered reverse transcriptase to reverse transcribe the RNA molecule to a complementary DNA (cDNA) molecule.
  • the contacting step may be in the presence of a plurality of nucleic acid barcode molecules, wherein each nucleic acid barcode molecule comprises a barcode sequence.
  • the nucleic acid barcode molecule may comprise a sequence configured to couple to a template RNA molecule.
  • RNA molecules include, without limitation, an oligo(dT) sequence, a random N-mer primer, or a target-specific primer.
  • the nucleic acid barcode molecule may comprise a template switching sequence. 111 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC [000413]
  • the RNA molecule is a messenger RNA (mRNA) molecule.
  • the contacting step provides conditions suitable to allow the engineered reverse transcriptase to: (i) transcribe the mRNA molecule into the cDNA molecule with the oligo(dT) sequence and/or (ii) perform a template switching reaction, thereby generating the cDNA molecule which comprises the barcode sequence, or a derivative thereof.
  • the contacting step may occur in (i) a partition having a reaction volume (e.g., as further described herein and see e.g., US Patent Nos.10,400,280 and 10,323,278, each of which is incorporated herein by reference in its entirety); (ii) in a bulk reaction where the reaction components (e.g., template RNA and engineered reverse transcriptase) are in solution; or (iii) on a nucleic acid array (see e.g., US Patent Nos.10,480,022 and 10,030,261 as well as WO/2020/047005 and WO/2020/047010, each of which is incorporated herein by reference in its entirety).
  • a reaction volume e.g., as further described herein and see e.g., US Patent Nos.10,400,280 and 10,323,278, each of which is incorporated herein by reference in its entirety
  • the reaction components e.g., template RNA and engineered reverse transcriptase
  • the reverse transcription reaction may occur in a tissue (e.g., in situ reverse transcription), on a template that is associated with a sequence on a substrate, such as practiced in spatial transcriptomics, or further in a RT-PCR or other reverse transcription reaction in vitro on a purified target, partially purified target or unpurified target as found for example in a cellular lysate.
  • tissue e.g., in situ reverse transcription
  • RT-PCR reverse transcription reaction in vitro on a purified target, partially purified target or unpurified target as found for example in a cellular lysate.
  • assays involving nucleic acid sample processing may include, but are not limited to, single-cell transcription profiling, single-cell sequence analysis, immune profiling of individual T and B cells, single-cell chromatin accessibility analysis (e.g., ATAC seq analysis), single cell processing and analysis, paired single cell TCR sequencing, paired TCR ⁇ and TCR ⁇ .
  • exemplary assays may be carried out using commercially available systems for encapsulating biological samples, gel beads, barcodes, and/or other compounds/materials in droplets, such as The Chromium System (10X Genomics, Pleasanton CA USA).
  • Engineered RT polypeptide or recombinant RT protein may be used in methods of profiling a T-Cell receptor (TCR).
  • TCR T-Cell receptor
  • the poly-dT sequence may be extended in a reverse transcription reaction using the mRNA as a template to produce a cDNA transcript complementary to the mRNA and also includes sequence of a barcode oligonucleotide.
  • Terminal transferase activity of the reverse transcriptase can add additional bases to the cDNA transcript (e.g., polyC).
  • the switch oligo may then hybridize with the additional bases added to the cDNA 112 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC transcript and facilitate template switching.
  • a sequence complementary to the switch oligo sequence can then be incorporated into the cDNA transcript via extension of the cDNA transcript using the switch oligo as a template.
  • all the cDNA transcripts of the individual mRNA molecules include a common barcode sequence.
  • the transcripts made from different mRNA molecules within a given partition will vary at this unique sequence. As described elsewhere herein, this provides a quantification feature that can be identifiable even following any subsequent amplification of the contents of a given partition, e.g., the number of unique segments associated with a common barcode can be indicative of the quantity of mRNA originating from a single partition, and thus, a single cell.
  • the cDNA transcript may then be amplified with PCR primers.
  • the amplified product may then be purified (e.g., via solid phase reversible immobilization (SPRI)).
  • SPRI solid phase reversible immobilization
  • the amplified product can be ligated to additional functional sequences, and further amplified (e.g., via PCR).
  • the functional sequences may include a sequencer specific flow cell attachment sequence such as but not limited to., a P7 sequence for Illumina ® sequencing systems, as well as functional sequence, which may include a sequencing primer binding site, e.g., for a R2 primer for Illumina ® sequencing systems, as well as functional sequence, which may include a sample index, e.g., an i7 sample index sequence for Illumina ® sequencing systems.
  • a sequencer specific flow cell attachment sequence such as but not limited to., a P7 sequence for Illumina ® sequencing systems, as well as functional sequence, which may include a sequencing primer binding site, e.g., for a R2 primer for Illumina ® sequencing systems, as well as functional sequence, which may include a sample index, e.g., an i7 sample index sequence for Illumina ® sequencing systems.
  • the present disclosure provides novel engineered reverse transcriptase 113 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC polypeptide or recombinant RT proteins that function efficiently in high throughput amplification reaction assays that require reaction volumes of less than about 1 nanoliter.
  • the method comprises providing a reaction volume which comprises an engineered reverse transcriptase and a template ribonucleic acid (RNA) molecule.
  • the contacting occurs in a reaction volume, which may be less than 1 nanoliter, less than 750 picoliters, or less than 500 picoliters.
  • the reaction volume is present in a partition, such as a droplet or well (including a microwell or a nanowell).
  • the engineered reverse transcriptases, the recombinant RT protein, or derivatives thereof as described herein are used in a reaction volume less than about 1 nanoliter (nL).
  • the engineered reverse transcriptases, the recombinant RT proteins, or derivatives thereof, as described herein are used in a reaction volume that is less than about 500 picoliter (pL).
  • the reaction volume is contained within a partition. In some embodiments, the reaction volume is contained within a droplet. In some embodiments, the reaction volume is contained within a droplet in an emulsion. In some embodiments, the reaction volume is contained within a droplet emulsion having a reaction volume of less than about 1 nL. In some embodiments, the reaction volume is contained within a droplet emulsion having a reaction volume of less than about 500 pL. [000421] In some embodiments, the reaction volume is contained within a well. In some embodiments, the reaction volume is contained within a well having a reaction volume less than about 1 nL. In some embodiments, the reaction volume is contained within a well.
  • the reaction volume is contained within a well having a reaction volume less than about 500 pL. In some embodiments, the reaction volume is contained within a well in an array of wells having an extracted nucleic acid molecule, and the template nucleic acid molecule is the extracted nucleic acid molecule. In some embodiments, the reaction volume is contained within a well in an array of wells having a cell comprising a template nucleic acid molecule, and where the template nucleic acid molecule is released from the cell. [000422] In another embodiment, a method comprises providing a reaction volume, which comprises an engineered reverse transcriptase and a template ribonucleic acid (RNA) molecule and is considered a “low volume reaction”.
  • RNA ribonucleic acid
  • the reaction volume may comprise a plurality of nucleic acid barcode molecules, and each nucleic acid barcode molecule comprises a barcode 114 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC sequence.
  • the contacting occurs in a reaction volume, a low volume reaction, which may be less than 1 nanoliter, less than 750 picoliters, or less than 500 picoliters.
  • the reaction volume is present in a partition, such as a droplet or well (including a microwell or a nanowell). 3.
  • the barcoding reaction produces single stranded complementary deoxyribonucleic acid (cDNA) molecules each having a molecular tag on a 5’ end thereof, followed by amplification of the cDNA to produce a double stranded DNA having the molecular tag on the 5’ end and a 3’ end of the double stranded DNA.
  • the molecular tags e.g., barcode oligonucleotides
  • the UMIs are oligonucleotides.
  • the molecular tags are coupled to priming sequences.
  • each of the priming sequences comprises a random N-mer sequence.
  • the random N-mer sequence is complementary to a 3’ sequence of the RNA molecules.
  • the priming sequence comprises a poly-dT sequence having a length of at least 5 bases.
  • the priming sequence comprises a poly-dT sequence having a length of at least 10 bases (SEQ ID NO: 4).
  • the priming sequence comprises a poly-dT sequence having a length of at least 5 bases, at least 6 bases, at least 7 bases, at least 8 bases, at least 9 bases, at least 10 bases.
  • UMIs Unique molecular identifiers
  • nucleic acid sequences are assigned or associated with individual cells or populations of cells, in order to tag or label the cell’s components (and as a result, its characteristics).
  • UMIs Unique molecular identifiers
  • These unique molecular identifiers may be used to attribute the cell’s components and characteristics to an individual cell or group of cells, additionally to be used as a method for counting the individual cells or groups of cells by their incorporation.
  • the unique molecular identifiers are provided in the form of nucleic acid molecules (e.g., oligonucleotides) that comprise nucleic acid barcode sequences that may be attached to or otherwise associated with the nucleic acid contents of individual cell, or to other components of the cell, and particularly to fragments of those nucleic acids.
  • the nucleic acid molecule can, and do have differing barcode sequences, or at least represent a large number of 115 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC different barcode sequences across all of the partitions in a given analysis.
  • nucleic acid barcode sequences can include from about 6 to about 20 or more nucleotides within the sequence of the nucleic acid molecules (e.g., oligonucleotides).
  • the nucleic acid barcode sequences can include from about 6 to about 20, 30, 40, 50, 60, 70, 80, 90, 100 or more nucleotides.
  • the length of a barcode sequence may be about 6, 7, 8, 9, 10, 11, 12, 13, 14, 15, 16, 17, 18, 19, 20 nucleotides or longer.
  • the length of a barcode sequence may be at least about 6, 7, 8, 9, 10, 11, 12, 13, 14, 15, 16, 17, 18, 19, 20 nucleotides or longer. In some cases, the length of a barcode sequence may be at most about 6, 7, 8, 9, 10, 11, 12, 13, 14, 15, 16, 17, 18, 19, 20 nucleotides or shorter. These nucleotides may be completely contiguous, i.e., in a single stretch of adjacent nucleotides, or they may be separated into two or more separate subsequences that are separated by 1 or more nucleotides. In some cases, separated barcode subsequences can be from about 4 to about 16 nucleotides in length.
  • the barcode subsequence may be about 4, 5, 6, 7, 8, 9, 10, 11, 12, 13, 14, 15, 16 nucleotides or longer. In some cases, the barcode subsequence may be at least about 4, 5, 6, 7, 8, 9, 10, 11, 12, 13, 14, 15, 16 nucleotides or longer. In some cases, the barcode subsequence may be at most about 4, 5, 6, 7, 8, 9, 10, 11, 12, 13, 14, 15, 16 nucleotides or shorter.
  • the resulting population of partitions can also include a diverse barcode library that may include at least about 1,000 different barcode sequences, at least about 5,000 different barcode sequences, at least about 10,000 different barcode sequences, at least at least about 50,000 different barcode sequences, at least about 100,000 different barcode sequences, at least about 1,000,000 different barcode sequences, at least about 5,000,000 different barcode sequences, or at least about 10,000,000 different barcode sequences.
  • each partition of the population can include at least about 1,000 nucleic acid molecules, at least about 5,000 nucleic acid molecules, at least about 10,000 nucleic acid molecules, at least about 50,000 nucleic acid molecules, at least about 100,000 nucleic acid molecules, at least about 500,000 nucleic acids, at least about 1,000,000 nucleic acid molecules, at least about 5,000,000 nucleic acid molecules, at least about 10,000,000 nucleic acid molecules, at least about 50,000,000 nucleic acid molecules, at least 116 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC about 100,000,000 nucleic acid molecules, at least about 250,000,000 nucleic acid molecules and in some cases at least about 1 billion nucleic acid molecules.
  • the enhanced reverse transcriptase activity of the engineered reverse transcriptase disclosed herein is an enhanced ability to yield mitochondrial UMI counts as compared to a reverse transcriptase having the amino acid sequence set forth in SEQ ID NO:1 or 15. In some embodiments, the enhanced reverse transcriptase activity is an enhanced ability to yield increased ribosomal UMI counts as compared to a reverse transcriptase having the amino acid sequence set forth in SEQ ID NO:1 or 15.
  • Read counting and UMI counting are the principal gene expression quantification schemes used in single-cell RNA-sequencing (scRNA- seq) analysis, as such with increased ribosomal UMI counts sensitivity and accuracy increases for a scRNA-seq assay in determining transcriptome profiles for any given cell, group of cells or tissues. Numerous metrics can be used for quality control of single-cell RNA-sequencing, including percent of reads mapping to ribosomal genes, percent of reads mapping to mitochondrial genes, total number of UMIs detected, or number of features to which 50% of the reads map.
  • the number of different UMIs can be indicative of the quantity of mRNA originating from a given partition, and thus from the cell.
  • the transcripts can be amplified, purified and sequenced to identify the sequence of the cDNA transcript of the mRNA, as well as to sequence the barcode segment and the UMI segment. While a poly-dT primer sequence is described, other targeted or random primer sequences may also be used in priming the reverse transcription reaction.
  • the nucleic acid molecules bound to the bead may be used to hybridize and capture the mRNA on the solid phase of the bead, for example, in order to facilitate the separation of the RNA from other cell contents.
  • certain reverse transcriptase enzymes may increase UMI reads from genes of a desired length or length of interest.
  • the desired length of genes may be selected from lengths comprising less than 500 nucleotides, between 500 and 1000 nucleotides, between 1000 and 1500 nucleotides and greater than 1500 nucleotides.
  • a reverse transcriptase may preferentially increase UMI reads from genes of one length range. It is 117 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC recognized that an engineered reverse transcriptase may perform similarly, differently or comparably in a 3’-reverse transcription assay or a 5’-reverse transcription assay. It is similarly recognized that an engineered reverse transcriptase may preferentially increase UMI reads from a length of genes in a 3’-reverse transcription assay than in a 5’-reverse transcription assay. 4.
  • the engineered reverse transcriptases or the recombinant RT protein of the present disclosure may be suitable for use in methods in which a cell can be co-partitioned along with a nucleic acid barcode molecule bearing bead.
  • the nucleic acid barcode molecules can be released from the bead in the partition.
  • the poly-dT poly-deoxythymine, also referred to as oligo (dT)
  • dT oligo
  • Reverse transcription results in a cDNA transcript of the mRNA, but that transcript includes each of the sequence segments of the nucleic acid molecule.
  • the nucleic acid molecule comprises an anchoring sequence, it may be more likely hybridize to and prime reverse transcription at the sequence end of the poly-A tail of the mRNA.
  • all of the cDNA transcripts of the individual mRNA molecules may include a common barcode sequence segment.
  • the transcripts made from the different mRNA molecules within a given partition may vary at the unique UMI segment.
  • the number of different UMIs can be indicative of the quantity of mRNA originating from a given partition, and thus from the cell.
  • the transcripts can be amplified, cleaned up and sequenced to identify the sequence of the cDNA transcript of the mRNA, as well as to sequence the barcode segment and the UMI segment. While a poly-dT primer sequence is described, other targeted or random priming sequences may also be used in priming the reverse transcription reaction.
  • the plurality of nucleic acid barcoded molecules are attached to a support (e.g., a particle, a slide, a chip, a bead, etc.).
  • the support is selected from an array, a bead, a gel bead, a microparticle, and a polymer.
  • the nucleic acid barcoded molecules attached to a support comprise molecular tags (UMIs), primer sequences, capture sequences, 118 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC cleavage sequences, or additional functional sequences.
  • UMIs molecular tags
  • the support is a gel bead.
  • the nucleic acid barcoded molecules are releasably attached to the gel bead.
  • the gel bead comprises a polyacrylamide polymer.
  • a cross-section of the gel bead is less than about 100 ⁇ m. In some embodiments, a cross-section of a gel bead is less than about 60 ⁇ m. In some embodiments, a cross-section of a gel bead is less than about 50 ⁇ m. In some embodiments, a cross-section of a gel bead is less than about 40 ⁇ m.
  • a cross-section of a gel bead is less than about 100 ⁇ m, less than about 99 ⁇ m, less than about 98 ⁇ m, less than about 97 ⁇ m, less than about 96 ⁇ m, less than about 95 ⁇ m, less than about 94 ⁇ m, less than about 93 ⁇ m, less than about 92 ⁇ m, less than about 91 ⁇ m, less than about 90 ⁇ m, less than about 89 ⁇ m, less than about 88 ⁇ m, less than about 87 ⁇ m, less than about 86 ⁇ m, less than about 85 ⁇ m, less than about 84 ⁇ m, less than about 83 ⁇ m, less than about 82 ⁇ m, less than about 81 ⁇ m, less than about 80 ⁇ m, less than about 79 ⁇ m, less than about 78 ⁇ m, less than about 77 ⁇ m, less than about 76 ⁇ m, less than about 75 ⁇ m, less than about 74 ⁇ m, less than about 82
  • nucleic acid molecules e.g., oligonucleotides
  • Functionalization of beads for attachment of nucleic acid molecules may be achieved through a wide range of different approaches, including activation of chemical groups within a polymer, incorporation of active or activatable functional groups in the polymer structure, or attachment at the pre-polymer or monomer stage in bead production.
  • precursors e.g., monomers, cross-linkers
  • precursors e.g., monomers, cross-linkers
  • precursors e.g., monomers, cross-linkers
  • bead may comprise acrydite moieties, such that when a bead is generated, the bead also comprises acrydite moieties.
  • the acrydite moieties can be attached to a nucleic acid molecule (e.g., oligonucleotide), which may include a priming sequence (e.g., a primer for amplifying target nucleic acids, random primer, primer sequence for messenger RNA) and/or one or more barcode sequences.
  • the one more barcode sequences may include sequences that are the same for all nucleic acid molecules coupled to a given bead and/or sequences that are different across 119 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC all nucleic acid molecules coupled to the given bead.
  • the nucleic acid molecule may be incorporated into the bead.
  • the nucleic acid molecule can comprise a functional sequence, for example, for attachment to a sequencing flow cell, such as, for example, a P5 sequence for Illumina ® sequencing.
  • the nucleic acid molecule or derivative thereof e.g., oligonucleotide or polynucleotide generated from the nucleic acid molecule
  • can comprise another functional sequence such as, for example, a P7 sequence for attachment to a sequencing flow cell for Illumina ® sequencing.
  • the nucleic acid molecule can comprise a barcode sequence.
  • the primer can comprise a unique molecular identifier (UMI).
  • the primer can comprise an R1 sequence for use in Illumina ® sequencing workflows. In some cases, the primer can comprise an R2 sequence for use in Illumina ® sequencing workflows.
  • nucleic acid molecules e.g., oligonucleotides, polynucleotides, etc.
  • examples of such nucleic acid molecules e.g., oligonucleotides, polynucleotides, etc.
  • uses thereof as may be used with compositions, devices, methods and systems of the present disclosure, are provided in U.S. Patent Pub. Nos.2014/0378345 and 2015/0376609, each of which is entirely incorporated herein by reference.
  • the present invention is not limited as to a composition of any nucleic acid molecule or derivative thereof, or any particular sequencing platform and these characterizations serve as examples only which may be useful in a reverse transcription workflow.
  • a cell in operation, can be co-partitioned along with a barcode bearing bead.
  • the barcoded nucleic acid molecules affixed to a bead can be released from the bead in the partition.
  • the poly-dT (poly-deoxythymine, also referred to as oligo (dT)) segment of one of the released nucleic acid molecules can hybridize to (e.g., capture)_the poly-A tail of a mRNA molecule.
  • Reverse transcription may result in a cDNA transcript of the mRNA which cDNA transcript also includes each of the sequence segments of the nucleic acid molecule.
  • the nucleic acid molecule comprises additional functional sequences (e.g., capture domains, primer domains, UMIs, barcodes, etc.), it can hybridize to and prime reverse transcription of the mRNA using the hybridized mRNA as a template.
  • additional functional sequences e.g., capture domains, primer domains, UMIs, barcodes, etc.
  • all of the cDNA transcripts of the individual mRNA molecules may include a common barcode sequence.
  • the transcripts made from the different mRNA molecules within a given partition may vary with respect to unique molecular 120 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC identifying sequences (e.g., UMIs).
  • the number of different UMIs can be indicative of the quantity of mRNA originating from a given partition, and thus from the cell.
  • the transcripts can be amplified and sequenced to identify the sequence of the original mRNA captured template, as well as the sequence of the associated barcode and UMI. While a poly-dT capture sequence is described, other targeted or random capture sequences may also be used in capture or hybridize to a template for initiating the reverse transcription reaction.
  • the poly-dT segment may be extended in a reverse transcription reaction using the mRNA as a template to produce a cDNA transcript complementary to the mRNA and also includes sequence segments of a barcode oligonucleotide.
  • Terminal transferase activity of the reverse transcriptase can add additional bases to the cDNA transcript (e.g., polyC).
  • the switch oligo may then hybridize with the additional bases added to the cDNA transcript and facilitate template switching.
  • a sequence complementary to the switch oligo sequence can then be incorporated into the cDNA transcript via extension of the cDNA transcript using the switch oligo as a template.
  • all the cDNA transcripts of the individual mRNA molecules include a common barcode sequence segment.
  • the transcripts made from different mRNA molecules within a given partition will vary at this unique sequence. As described elsewhere herein, this provides a quantification feature that can be identifiable even following any subsequent amplification of the contents of a given partition, e.g., the number of unique segments associated with a common barcode can be indicative of the quantity of mRNA originating from a single partition, and thus, a single cell.
  • the cDNA transcript may then be amplified with PCR primers.
  • the amplified product may then be purified (e.g., via solid phase reversible immobilization (SPRI)).
  • the amplified product may be sheared, ligated to additional functional sequences, and further amplified (e.g., via PCR).
  • Any of the engineered RT enzymes of the present disclosure including without limitation any of the enzymes comprising the amino acid sequence and/or non-limiting embodiment of nucleic acid sequences shown in Table 1, or Table 2, could be analyzed in any suitable assay, including without limitation the assays described herein.
  • Assays include without limitation 5’ gene expression analyses, with or without VDJ analysis, 3’ gene expression 121 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC analysis, epigenetic analysis, or multiomic analyses.
  • the sample can comprise a cell, a cell bead, a permeabilized cell, or a nucleus.
  • the cell can be fixed.
  • the cell can be permeabilized.
  • the cell can be permeabilized and fixed.
  • the cell bead is fixed.
  • the nucleus is permeabilized or fixed.
  • the nucleus is permeabilized and fixed.
  • the biological sample comprises a suitable cellular preparation selected from cell populations and/or single cells.
  • the biological sample comprises a tissue.
  • the sample can also be a suitable cellular preparation selected from cell populations and/or single cells.
  • the sample can be a tissue.
  • the sample comprises cells in suspension, fresh cells, or fixed cells.
  • the sample can also comprise cells and tissues immobilized on various solid surfaces.
  • the reverse transcription reaction described herein is part of a single cell RNA sequencing assay.
  • the single cell RNA sequencing assay further comprises, prior to the reverse transcription, partitioning the cell, the cell bead, or the nucleus into a partition. In one embodiment, the single cell RNA sequencing assay further comprises, after the reverse transcription reaction, hybridizing the nucleic acid product to an oligonucleotide molecule comprising a partition-specific barcode.
  • the reverse transcription reaction described herein is part of a spatial RNA sequencing assay.
  • the sample is a fresh tissue. In some embodiments, the sample is a frozen sample. In some embodiments, the sample was previously frozen.
  • the sample is a formalin-fixed, or paraffin embedded (FFPE) sample.
  • FFPE formalin-fixed, or paraffin embedded
  • 10X Genomics Ref.: 100-165501PC samples generally are heavily cross-linked and fragmented, and therefore this type of sample allows for limited RNA recovery using conventional detection techniques.
  • methods of targeted RNA capture provided herein are less affected by RNA degradation associated with FFPE fixation than other methods (e.g., methods that take advantage of oligo-dT capture and reverse transcription of mRNA).
  • methods provided herein enable sensitive measurement of specific genes of interest that otherwise might be missed with a whole transcriptomic approach.
  • a biological sample e.g., tissue section
  • methanol stained with hematoxylin and eosin
  • fixing, staining, and imaging occurs before one or more oligonucleotide probes are hybridized to the sample.
  • a destaining step e.g., a hematoxylin and eosin destaining step
  • destaining can be performed by performing one or more (e.g., one, two, three, four, or five) washing steps (e.g., one or more (e.g., one, two, three, four, or five) washing steps performed using a buffer including HCl).
  • the images can be used to map spatial gene expression patterns back to the biological sample.
  • a permeabilization enzyme can be used to permeabilize the biological sample directly on the slide.
  • the methods of targeted RNA capture as disclosed herein include hybridization of multiple probe oligonucleotides. In some embodiments, the methods include 2, 3, 4, or more probe oligonucleotides that hybridize to one or more analytes of interest.
  • the methods include two probe oligonucleotides.
  • the probe oligonucleotide includes sequences complementary that are complementary or substantially complementary to an analyte.
  • the probe oligonucleotide includes a sequence that is complementary or substantially complementary to an analyte (e.g., an mRNA of interest (e.g., to a portion of the sequence of an mRNA of interest)).
  • analyte e.g., an mRNA of interest (e.g., to a portion of the sequence of an mRNA of interest)
  • a method of analyzing a sample comprising a nucleic acid molecule may comprise providing a plurality of nucleic acid molecules (e.g., RNA molecules), where each nucleic acid molecule comprises a first target region (e.g., a sequence that is 3′ of a target sequence or a sequence that is 5′ of a target sequence) and a second target region (e.g., a 123 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC sequence that is 5′ of a target sequence or a sequence that is 3′ of a target sequence), a plurality of first probe oligonucleotides, and a plurality of second probe oligonucleotides.
  • a first target region e.g., a sequence that is 3′ of a target sequence or a sequence that is 5′ of a target sequence
  • a second target region e.g., a 123 4876-6828-
  • the templated ligation methods that allow for targeted RNA capture as provided herein include a first probe oligonucleotide and a second probe oligonucleotide.
  • the first and second probe oligonucleotides each include sequences that are substantially complementary to the sequence of an analyte of interest.
  • the first and/or second probe oligonucleotide is at least 70%, at least 75%, at least 80%, at least 85%, at least 90%, at least 91%, at least 92%, at least 93%, at least 94%, at least 95%, at least 96%, at least 97%, at least 98%, at least 99%, or 100% complementary to a sequence in an analyte.
  • the first probe oligonucleotide and the second probe oligonucleotide hybridize to adjacent sequences on an analyte.
  • the first and/or second probe as disclosed herein includes one of at least two ribonucleic acid bases at the 3′ end; a functional sequence; a phosphorylated nucleotide at the 5′ end; and/or a capture probe binding domain.
  • the functional sequence is a primer sequence.
  • the capture probe binding domain is a sequence that is complementary to a particular capture domain present in a capture probe.
  • the capture probe binding domain includes a poly(A) sequence.
  • the capture probe binding domain includes a poly-uridine sequence, a poly-thymidine sequence, or both.
  • the capture probe binding domain includes a random sequence (e.g., a random hexamer or octamer). In some embodiments, the capture probe binding domain is complementary to a capture domain in a capture probe that detects a particular target(s) of interest. [000451] In some embodiments, a capture probe binding domain blocking moiety that interacts with the capture probe binding domain is provided. In some instances, the capture probe binding domain blocking moiety includes a nucleic acid sequence. In some instances, the capture probe binding domain blocking moiety is a DNA oligonucleotide. In some instances, the capture probe binding domain blocking moiety is an RNA oligonucleotide.
  • a capture probe binding domain blocking moiety includes a sequence that is complementary or substantially complementary to a capture probe binding domain. In some embodiments, a capture probe binding domain blocking moiety prevents the capture probe binding domain from binding 124 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC the capture probe when present. In some embodiments, a capture probe binding domain blocking moiety is removed prior to binding the capture probe binding domain (e.g., present in a ligated probe) to a capture probe. In some embodiments, a capture probe binding domain blocking moiety comprises a poly-uridine sequence, a poly-thymidine sequence, or both.
  • the first probe oligonucleotide hybridizes to an analyte.
  • the second probe oligonucleotide hybridizes to an analyte.
  • both the first probe oligonucleotide and the second probe oligonucleotide hybridize to an analyte. Hybridization can occur at a target having a sequence that is 100% complementary to the probe oligonucleotide(s).
  • hybridization can occur at a target having a sequence that is at least (e.g., at least about) 80%, at least (e.g., at least about) 85%, at least (e.g., at least about) 90%, at least (e.g., at least about) 95%, at least (e.g., at least about) 96%, at least (e.g., at least about) 97%, at least (e.g., at least about) 98%, or at least (e.g., at least about) 99% complementary to the probe oligonucleotide(s).
  • the first probe oligonucleotide is extended.
  • the second probe oligonucleotide is extended. Extending probes can be accomplished using any method disclosed herein.
  • a polymerase e.g., a DNA polymerase
  • methods disclosed herein include a wash step. In some instances, the wash step occurs after hybridizing the first and the second probe oligonucleotides.
  • the wash step removes any unbound oligonucleotides and can be performed using any technique or solution disclosed herein or known in the art. In some embodiments, multiple wash steps are performed to remove unbound oligonucleotides.
  • the probe oligonucleotides e.g., first and the second probe oligonucleotides
  • the probe oligonucleotides are ligated together, creating a single ligated probe that is complementary to the analyte. Ligation can be performed enzymatically or chemically, as described herein.
  • the engineered RT polypeptide or the recombinant RT protein enhances template switching (TS) efficiency, processivity efficiency, binding affinity, transcription efficiency, chemical tolerance, ability to yield mitochondrial unique molecular identity (UMI) counts, ability to yield ribosomal unique molecular identity (UMI) counts, strand displacement, end-to-end template jumping, or any combination thereof, when compared to an RT polypeptide or a recombinant RT protein lacking a conjugated DNA binding domain.
  • TS template switching
  • UMI mitochondrial unique molecular identity
  • UMI ribosomal unique molecular identity
  • the engineered RT polypeptide or the recombinant RT protein enhances at least two or more of template switching (TS) efficiency, processivity efficiency, binding affinity, transcription efficiency, chemical tolerance, ability to yield mitochondrial unique molecular identity (UMI) counts, strand displacement, end-to-end template jumping, or ability to yield ribosomal unique molecular identity (UMI) counts, when compared to an RT polypeptide or a recombinant RT protein lacking a conjugated DNA binding domain.
  • TS template switching
  • UMI mitochondrial unique molecular identity
  • UMI ribosomal unique molecular identity
  • the recombinant RT protein or the engineered RT enhances transcript capture during amplification, when compared to an RT polypeptide or a recombinant RT protein lacking a conjugated DNA binding domain.
  • the DNA binding domain of the recombinant RT protein or the engineered RT enhances the hybridization of a transcript and a primer during the amplification process, when compared to an RT polypeptide or a recombinant RT protein lacking a conjugated DNA binding domain.
  • Engineered reverse transcriptase polypeptides or recombinant RT proteins described herein may be used in methods of a T-Cell receptor (TCR) and a B-cell receptor (BRC) profiling.
  • TCR T-Cell receptor
  • BRC B-cell receptor
  • an engineered reverse transcriptase is used in methods including but not limited to processing of a TCR from an individual T cell(s) or groups of T cell(s), determining the nucleotide sequence of the TCR(s) of T cell(s), and obtaining TCR repertoire profile.
  • a nucleic acid barcode sequence is appended to a nucleic acid molecule encoding for a TCR (e.g.
  • a TCR such as a TCR ⁇ and/or a TCR ⁇ mRNA
  • a barcoded nucleic acid molecule may serve as a template, such as a template polynucleotide, that can be further processed (e.g., amplified) and sequenced to obtain the target nucleic acid sequence.
  • a barcoded nucleic acid molecule may be further processed (e.g., amplified) and sequenced to obtain the nucleic acid sequence of the TCR.
  • TCR is a molecule found on the surface of T cells. Typically binding of the TCR by an antigenic molecule results in cell activation and response. The TCR is a heterodimer composed of two different protein chains. In many T cells, these two proteins are alpha ( ⁇ ) and beta ( ⁇ ) chains.
  • T cells these two proteins are gamma ( ⁇ ) and delta ( ⁇ ) chains.
  • the ratio of TCRs comprised of ⁇ / ⁇ chains versus ⁇ / ⁇ chains may change during a diseased state such as cancer, tumor, infectious disease, inflammatory disease or autoimmune disease.
  • Engagement of the TCR with a peptide-MHC activates a T cell through a series of biochemical events mediated by associated enzymes, co-receptors, specialized adaptor molecules, and activated or released transcription factors.
  • Each of the two chains of a TCR contains multiple copies of gene segments- a variable ‘V’ gene segment, a diversity ‘D’ segment and a joining ‘J’ segment.
  • the TCR alpha chain is generated by recombination of V and J segments, while the beta chain is generated by recombination of V, D and J segments.
  • generation of the TCR gamma chain involves recombination of V and J segments.
  • Generation of the TCR delta chain occurs by recombination of V, D and J gene segments. The intersection of these specific regions (V and J for the alpha or gamma chain, or V,D, J for the beta or delta chain) corresponds to the CDR3 region involved in antigen-MHC recognition.
  • Complementarity determining regions e.g., CDR1, CDR2 and CDR3 or hypervariable regions are sequences in the variable domains of antigen receptors (e.g., T cell receptor and immunoglobulin) that can complement an antigen.
  • antigen receptors e.g., T cell receptor and immunoglobulin
  • Most of the diversity of CDRs is found in CDR3, with the diversity being generated by somatic recombination events during the development of T lymphocytes.
  • CDR3 which is encoded by the junctional region 127 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC between the V and J or D and J genes, is highly variable.
  • CDR3 is often used as a region of interest to determine T cell clonotypes, a unique nucleotide sequence that arises during the gene rearrangement process, as it is highly unlikely that two T cells will express the same CDR3 nucleotide sequence unless they are derived from the same clonally expanded T cell.
  • TCR gene sequences may include, but are not limited to, sequences of various T cell receptor alpha variable genes (TRAV genes), T cell receptor alpha joining genes (TRAJ genes), T cell receptor alpha constant genes (TRAC genes), T cell receptor beta variable genes (TRBV genes), T cell receptor beta diversity genes (TRBD genes), T cell receptor beta joining genes (TRBJ genes), T cell receptor gamma variable genes (TRGV genes), T cell receptor gamma joining genes (TRGJ genes), T cell receptor gamma constant genes (TRGC genes), T cell receptor delta variable genes (TRDV genes), T cell receptor delta diversity genes (TRDD genes), T cell receptor delta joining genes (TRDJ genes) and T cell receptor delta constant genes (TRDC genes).
  • TRAV genes T cell receptor alpha variable genes
  • TRAJ genes T cell receptor alpha joining genes
  • TRBV genes T cell receptor beta variable genes
  • TRBD genes T cell receptor beta diversity genes
  • TRBJ genes T cell receptor beta joining genes
  • TRGV genes T cell receptor gamma variable genes
  • kits comprising the engineered reverse transcriptase polypeptide or recombinant RT protein, the DNA binding domains or a derivative thereof as described herein.
  • the kit comprises one or more of a vector, a nucleotide, a buffer, a composition, a salt, and/or instructions.
  • a kit may comprise an engineered reverse transcriptase polypeptide or recombinant RT protein or a derivative thereof for use in reverse transcription or amplification of a nucleic acid molecule.
  • a kit may be used for single cell profiling of the transcriptome.
  • a kit may be used for spatial transcriptomics methods and assays.
  • a kit may be used for in situ methods and assays.
  • the kit may include suitable reaction buffers, dNTPs, one or more primers, one or more control reagents, or any other reagents disclosed for performing the methods of the present disclosure.
  • the engineered reverse transcriptase polypeptide or recombinant RT protein or a derivative thereof, reaction buffer, and dNTPs may be provided separately or may be provided together in a master mix solution.
  • the master mix is present at a concentration at least two times the working concentration indicated in instructions for use in an extension reaction. In other cases, the master mix may be present at a concentration at least three times, at least four times, at least five times, at least six times, at least seven times, at least eight times, at least nine times, or at least ten times, the working concentration indicated.
  • the primer in the kits may be a poly-dT primer, a random N-mer primer, or a target-specific primer.
  • the kits may further include one, two, three, four, five or more, up to all of partitioning fluids, including both aqueous buffers and non-aqueous partitioning fluids or oils, nucleic acid barcode capture probes that are releasably associated with beads, as described herein, microfluidic devices, reagents for disrupting cells, reagents for amplifying nucleic acids, as well as instructions for using any of the foregoing in the methods described herein.
  • the instructions for using any of the methods are generally recorded on a suitable recording medium (e.g., printed on a substrate such as paper or plastic), or available in a digital format.
  • the instructions may be present in the kits as a package insert, in the labeling of the container of the kit or components thereof (i.e., associated with the packaging or subpackaging).
  • the instructions may be present as an electronic storage data file present on a suitable computer readable storage medium.
  • the actual instructions may not be present in the kit but means for obtaining the instructions from a remote source, e.g., via the internet, may be provided.
  • Kits according to this aspect of the present disclosure comprise a carrier means, such as a box, carton, tube or the like, having in close confinement therein one or more container means, such as vials, tubes, ampoules, bottles and the like, wherein a first container means contains one or more of the engineered reverse transcriptase polypeptide or recombinant RT proteins or derivatives thereof of the present disclosure having reverse transcriptase activity.
  • kits of the 129 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC disclosure can also comprise (in the same or separate containers) one or more DNA polymerases, a suitable buffer, one or more nucleotides and/or one or more primers.
  • the kits of the disclosure can also comprise one or more hosts or cells including those that are competent to take up nucleic acids (e.g., DNA molecules including vectors).
  • Preferred hosts may include chemically competent or electrocompetent bacteria such as E. coli (including DH5, DH5 ⁇ , DH10B, HB101, Top 10, and other K-12 strains as well as E. coli B and E. coli W strains).
  • E. coli including DH5, DH5 ⁇ , DH10B, HB101, Top 10, and other K-12 strains as well as E. coli B and E. coli W strains).
  • kits of the disclosure can include one or more components (in mixtures or separately) including one or more engineered reverse transcriptase polypeptides or recombinant RT proteins or derivatives thereof having reverse transcriptase activity of the disclosure, one or more nucleotides (one or more of which may be labeled, e.g., fluorescently labeled) used for synthesis of a nucleic acid molecule, and/or one or more primers (e.g., oligo(dT) for reverse transcription, randomers for extension reactions, etc.).
  • Such kits can comprise one or more DNA polymerases.
  • the near or approximating unrecited number may be a number which, in the context in which it is presented, provides the substantial equivalent of the specifically recited number. If the degree of approximation is not otherwise clear from the context, “about” means either within plus or minus 10% of the provided value, or rounded to the nearest significant figure, in all cases inclusive of the provided value. In some embodiments, the term “about” indicates the designated value ⁇ up to 10%, up to ⁇ 5%, or up to ⁇ 1%.
  • the term “about” or “approximately” as used herein means within an acceptable error range for the particular value as determined by one of ordinary skill in the art, which will depend in part on how the value is measured or determined, i.e., the limitations of the measurement system.
  • “about” can mean within an acceptable standard deviation, per the practice in the art.
  • “about” can mean a range of up to ⁇ 20%, preferably up to ⁇ 10%, more preferably up to ⁇ 5%, and more preferably still up to ⁇ 1% of a given value.
  • the term can mean within an order of magnitude, preferably within 2-fold, of a value.
  • Analyte is intended a biological molecule. Analytes include but are not limited to a DNA analyte, an RNA analyte, an oligonucleotide, a reporter molecule, a reporter molecule configured to directly couple to a protein, a reporter molecule configured to indirectly couple to a protein, a reporter molecule configured to directly couple to a metabolite, and a reporter molecule configured to indirectly couple to a metabolite. [000482]
  • Adaptor(s),” “Adapter(s)” and “Tag(s)” may be used synonymously.
  • Barcoded nucleic acid molecule generally refers to a nucleic acid molecule that results from, for example, the processing of a nucleic acid barcoded molecule with a nucleic acid sequence (e.g., nucleic acid sequence complementary to a nucleic acid primer sequence encompassed by the nucleic acid barcoded molecule).
  • the nucleic acid sequence may be a targeted sequence or a non-targeted sequence.
  • the nucleic acid barcoded 132 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC molecule may be coupled to or attached to the nucleic acid molecule comprising the nucleic acid sequence.
  • a nucleic acid barcoded molecule described herein may be hybridized to an analyte (e.g., a messenger RNA (mRNA) molecule) of a cell. Reverse transcription can generate a barcoded nucleic acid molecule that has a sequence corresponding to the nucleic acid sequence of the mRNA and the barcode sequence (or a reverse complement thereof).
  • analyte e.g., a messenger RNA (mRNA) molecule
  • Reverse transcription can generate a barcoded nucleic acid molecule that has a sequence corresponding to the nucleic acid sequence of the mRNA and the barcode sequence (or a reverse complement thereof).
  • the processing of the nucleic acid molecule comprising the nucleic acid sequence, the nucleic acid barcoded molecule, or both can include a nucleic acid reaction, such as, in non-limiting examples, reverse transcription, nucleic acid extension, ligation, etc.
  • the nucleic acid reaction may be performed prior to, during, or following barcoding of the nucleic acid sequence to generate the barcoded nucleic acid molecule.
  • the nucleic acid molecule comprising the nucleic acid sequence may be subjected to reverse transcription and then be attached to the nucleic acid barcoded molecule to generate the barcoded nucleic acid molecule, or the nucleic acid molecule comprising the nucleic acid sequence may be attached to the nucleic acid barcoded molecule and subjected to a nucleic acid reaction (e.g., extension, ligation) to generate the barcoded nucleic acid molecule.
  • a barcoded nucleic acid molecule may serve as a template, such as a template polynucleotide, that can be further processed (e.g., amplified) and sequenced to obtain the target nucleic acid sequence.
  • a barcoded nucleic acid molecule may be further processed (e.g., amplified) and sequenced to obtain the nucleic acid sequence of the nucleic acid molecule (e.g., mRNA).
  • a nucleic acid barcoded molecule of a plurality of nucleic acid molecules may be used to generate a “barcoded nucleic acid molecule.”
  • a barcoded molecule comprises a different reporter barcode sequence that identifies a second analyte.
  • a different reporter barcode sequence or an analyte-specific barcode sequence may identify a protein, a lipid, a metabolite or other second analyte.
  • Barcoded nucleic acids may be generated (e.g., via a nucleic acid reaction, such as nucleic acid extension or ligation) from the constructs described in FIG.12.
  • capture handle sequence may then be hybridized to complementary sequence, such as capture sequence 1223 to generate (e.g., via a nucleic acid reaction, such as nucleic acid extension or ligation) a barcoded nucleic acid molecule comprising cell (e.g., partition specific) barcode sequence 1222 (or a reverse complement thereof) and reporter barcode sequence 1222 (or a 133 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC reverse complement thereof).
  • capture handle sequence 4323 comprises a sequence complementary to a template switching oligonucleotide on the capture sequence 1223.
  • the nucleic acid barcoded molecule 1290 e.g., partition-specific barcoded molecule
  • UMI not shown
  • Barcoded nucleic acid molecules can then be optionally processed as described elsewhere herein, e.g., to amplify the molecules and/or append sequencing platform specific sequences to the fragments. See, e.g., U.S. Pat. Pub.2018/0105808, which is hereby entirely incorporated by reference for all purposes. Barcoded nucleic acid molecules, or derivatives generated therefrom, can then be sequenced on a suitable sequencing platform.
  • analysis of multiple analytes may be performed.
  • analysis of an analyte e.g., a nucleic acid, a polypeptide, a carbohydrate, a lipid, a glycan, a glycan motif, a metabolite, a protein, etc.
  • an analyte e.g., a nucleic acid, a polypeptide, a carbohydrate, a lipid, a glycan, a glycan motif, a metabolite, a protein, etc.
  • a nucleic acid barcoded molecule 1290 e.g., partition specific barcoded molecule
  • nucleic acid barcoded molecule 1290 is attached to a support 1230 (e.g., a bead, such as a gel bead), such as those described elsewhere herein.
  • a support 1230 e.g., a bead, such as a gel bead
  • nucleic acid barcoded molecule 1290 may be attached to support 1230 via a releasable linkage 1240 (e.g., comprising a labile bond), such as those described elsewhere herein.
  • Nucleic acid barcoded molecule 1290 may comprise a functional sequence 1221 and optionally comprise other additional sequences, for example, a barcode sequence 1222 (e.g., common barcode, partition-specific barcode, or other functional sequences described elsewhere herein), and/or a UMI sequence (not shown).
  • the nucleic acid barcoded molecule 1290 may comprise a capture sequence 1223 that may be complementary to another nucleic acid sequence, such that it may hybridize to a particular sequence, e.g., capture handle sequence 1223.
  • capture sequence 1223 may comprise a poly-T sequence and may be used to hybridize to mRNA.
  • nucleic acid barcoded molecule 1290 comprises capture sequence 1223 complementary to a sequence of RNA molecule 1260 from a cell.
  • capture sequence 1223 comprises a sequence specific for an RNA molecule.
  • Capture sequence 1223 may comprise a known or targeted sequence or a random sequence.
  • a nucleic acid extension reaction may be 134 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC performed, thereby generating a barcoded nucleic acid product comprising capture sequence 12323, the functional sequence 1221, barcode sequence 1222, any other functional sequence, and a sequence corresponding to the RNA molecule 1260.
  • capture sequence 1223 may be complementary to an overhang sequence or an adapter sequence that has been appended to an analyte. Any suitable agent may degrade beads. Suitable agents may include, but are not limited to, changes in temperature, changes in pH, reduction, oxidation and exposure to water or other aqueous solutions.
  • a cell that is bound to labelling agent which is conjugated to oligonucleotide and support 1230 e.g., a bead, such as a gel bead
  • oligonucleotide and support 1230 e.g., a bead, such as a gel bead
  • nucleic acid barcoded molecule 1290 is partitioned into a partition amongst a plurality of partitions (e.g., a droplet of a droplet emulsion, a well of a microwell array, a fixed cell and/or nucleus, a fixed and permeabilized cell and/or nucleus).
  • the term “Bead,” as used herein, generally refers to a particle. The bead may be a solid or semi-solid particle.
  • the bead may be a gel bead.
  • the gel bead may include a polymer matrix (e.g., matrix formed by polymerization or cross-linking).
  • the polymer matrix may include one or more polymers (e.g., polymers having different functional groups or repeat units). Polymers in the polymer matrix may be randomly arranged, such as in random copolymers, and/or have ordered structures, such as in block copolymers. Cross-linking can be via covalent, ionic, or inductive, interactions, or physical entanglement.
  • the bead may be a macromolecule.
  • the bead may be formed of nucleic acid molecules bound together.
  • the bead may be formed via covalent or non-covalent assembly of molecules (e.g., macromolecules), such as monomers or polymers. Such polymers or monomers may be natural or synthetic. Such polymers or monomers may be or include, for example, nucleic acid molecules (e.g., DNA or RNA).
  • the bead may be formed of a polymeric material.
  • the bead may be magnetic or non-magnetic.
  • the bead may be rigid.
  • the bead may be flexible and/or compressible.
  • the bead may be disruptable or dissolvable.
  • the bead may be a solid particle (e.g., a metal-based particle including but not limited to iron oxide, gold or silver) covered with a coating comprising one or more polymers.
  • each when used in reference to a collection of items, is intended to identify an individual item in the collection but does not necessarily refer to every 135 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC item in the collection, unless expressly stated otherwise, or unless the context of the usage clearly indicates otherwise.
  • Efficiency in the context of a nucleic acid modifying enzyme of this invention refers to the ability of the enzyme to perform its catalytic function under specific reaction conditions. Typically, “efficiency” as defined herein is indicated by the amount of product generated under given reaction conditions.
  • the term “Enhances” in the context of an enzyme refers to improving the activity of the enzyme, i.e., increasing the amount of product per unit enzyme per unit time.
  • the term “Fidelity” refers to the accuracy of polymerization, or the ability of the reverse transcriptase to discriminate correct from incorrect substrates, (e.g., nucleotides) when synthesizing nucleic acid molecules which are complementary to a template.
  • % homology refers to the level of nucleic acid or amino acid sequence identity between the nucleic acid sequence that encodes any one of the inventive polypeptides (e.g., variant reverse transcriptases) or the inventive polypeptide's amino acid sequence, when aligned using a sequence alignment program.
  • the term “Identical” in the context of two nucleic acids or polypeptide sequences refers to the residues in the two sequences that are the same when aligned for maximum correspondence, as measured using a sequence comparison algorithms. Sequence comparison algorithms are known to those skill in the art. See. e.g., ebi.ac.uk/Tools/msa/clustalo/. [000497] As used herein, the term “Inhibitor resistance” refers to the ability of a reverse transcriptase to perform reverse transcription in the presence of a compound, chemical, protein, buffer, etc. that is typically inhibitory to the reverse transcriptase (prevents or inhibits reverse transcriptase activity).
  • the term “Low volume reaction” means a reaction volume less than 1 nanoliter, less than 750 picoliters, or less than 500 picoliters.
  • the term “Molecular tag,” as used herein, generally refers to a molecule capable of binding to a macromolecular constituent.
  • the molecular tag may bind to the macromolecular constituent with high affinity.
  • the molecular tag may bind to the macromolecular constituent with high specificity.
  • the molecular tag may comprise a nucleotide sequence.
  • the molecular tag may comprise a nucleic acid sequence.
  • the nucleic acid sequence may be at least a portion or an entirety of the molecular tag.
  • the molecular tag may be a nucleic acid molecule or may be part of a nucleic acid molecule.
  • the molecular tag may be an oligonucleotide or a polypeptide.
  • the molecular tag may comprise a DNA aptamer.
  • the molecular tag may be or comprise a primer.
  • the molecular tag may be, or comprise, a protein.
  • the molecular tag may comprise a polypeptide.
  • the molecular tag may be a barcode. [000500]
  • the term “mutation” or “mutant” or “variant“ indicates a change or changes introduced in a wild-type DNA sequence or a wildtype amino acid sequence.
  • mutations or variants include, but are not limited to, substitutions, insertions, deletions, and point mutations. Mutations can be made either at the nucleic acid level or at the amino acid level.
  • the term “Operably linked” or “conjugated” or “fusion” means that, in relation to the engineered RT polypeptide or the recombinant RT protein sequence, there are one or more sequences at the N or C terminus that, when transcribed and translated, create additional polypeptides in association with the enzyme amino acid sequence, thereby created a conjugation or fusion of one or more polypeptides from one expression vector.
  • Partition refers to a space or volume that may be suitable to contain one or more species or conduct one or more reactions.
  • a partition may be a physical compartment, such as a droplet, well or a fixed and/or permeabilized cell and/or nucleus.
  • the partition may isolate space or volume from another space or volume.
  • the droplet may be a first phase (e.g., aqueous phase) in a second phase (e.g., oil) immiscible with the first phase.
  • the droplet may be a first phase in a second phase that does not phase separate from the first phase, such as, for example, a capsule or liposome in an aqueous phase.
  • a partition may comprise one or more other (inner) partitions.
  • a partition may be a virtual compartment that can be defined and identified by an index (e.g., indexed libraries) across 137 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC multiple and/or remote physical compartments.
  • a physical compartment may comprise a plurality of virtual compartments.
  • partitions systems and methods for partitioning of one or more particles (such as, but not limited to, biological particles, macromolecular constituents of biological particles, beads, reagents, etc.) into discrete compartments or partitions (referred to interchangeably here as partitions), wherein each partition maintains separation of its own content from the contents of other partitions are known in the art. See for example US 2020/0032335, herein incorporated by reference in its entirety.
  • the partition can be a droplet in an emulsion.
  • a partition may comprise one or more other partitions.
  • a “plurality of nucleic acid barcoded molecules” may comprise at least about 500 nucleic acid barcoded molecules, at least about 1,000 nucleic acid barcoded molecules, at least about 5,000 nucleic acid barcoded molecules, at least about 10,000 nucleic acid barcoded molecules, at least about 50,000 nucleic acid barcoded molecules, at least about 100,000 nucleic acid barcoded molecules, at least about 500,000 nucleic acid barcoded molecules, at least about 1,000,000 barcoded molecules, at least about 5,000,000 nucleic acid barcoded molecules, at least about 10,000,000 nucleic acid barcoded molecules, at least about 100,000,000 nucleic acid barcoded molecules, at least about 1,000,000,000 nucleic acid barcoded molecules.
  • a plurality of nucleic acid barcoded molecules comprise a partition-specific barcode sequence.
  • Each of the plurality of nucleic acid barcoded molecules may include an identifier sequence separate from the partition-specific barcode sequence, where the identifier sequence is different for each nucleic acid partition-specific barcoded molecule of the plurality of nucleic acid partition specific barcoded molecules.
  • an identifier sequence is a unique molecular identifier (UMI) as described elsewhere herein.
  • UMI sequences can uniquely identify a particular nucleic acid molecule that is barcoded, which may be identifying particular nucleic acid molecules that are analyzed, counting particular nucleic acid molecules that are analyzed, etc.
  • each of the plurality of nucleic acid barcoded molecules can comprise the partition specific barcode sequence and the bead can 138 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC be from plurality of beads, such as a population of barcoded beads.
  • Each of the partition specific barcode sequences can be different from partition specific barcode sequences of nucleic acid barcoded molecules of other beads of the plurality of beads. Where this is the case, a population of barcoded beads, with each bead comprising a different partition specific barcode sequence can be analyzed.
  • the term “Processivity” refers to the ability of a reverse transcriptase to continuously extend a primer without disassociating from the nucleic acid template.
  • the length of a template a reverse transcriptase or polymerase is capable of replicating can also be used to describe the processivity of that reverse transcriptase or polymerase.
  • “Processivity” refers to the ability of a polymerase to remain bound to the template or substrate and perform DNA synthesis. Processivity is measured by the number of catalytic events that take place per binding event.
  • Reverse transcriptase activity indicates the capability of an enzyme to synthesize a DNA strand (that is, complementary DNA or cDNA) using RNA as a template. Reverse transcriptase activity may be measured by incubating an enzyme in the presence of an RNA template and deoxynucleotides, in the presence of an appropriate buffer, under appropriate conditions, for example as described in the Example below.
  • RT activity comprises the engineered RT fusion protein described herein or the engineered RT variant described herein.
  • Reverse transcriptase RT is used in its broadest sense to refer to any enzyme that exhibits reverse transcription activity as measured by methods disclosed herein or known in the art.
  • a "reverse transcriptase” of the present invention therefore, includes reverse transcriptases from retroviruses, other viruses, as well as a DNA polymerase exhibiting 139 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC reverse transcriptase activity, such as Tth DNA polymerase, Taq DNA polymerase, Tne DNA polymerase, Tma DNA polymerase, etc.
  • RT from retroviruses include, but are not limited to, Moloney Murine Leukemia Virus (M-MLV) RT, Human Immunodeficiency Virus (HIV) RT, Avian Sarcoma-Leukosis Virus (ASLV) RT, Rous Sarcoma Virus (RSV) RT, Avian Myeloblastosis Virus (AMV) RT, Avian Erythroblastosis Virus (AEV) Helper Virus MCAV RT, Avian Myelocytomatosis Virus MC29 Helper Virus MCAV RT, Avian Reticuloendotheliosis Virus (REV-T) Helper Virus REV-A RT, Avian Sarcoma Virus UR2 Helper Virus UR2AV RT, Avian Sarcoma Virus Y73 Helper Virus YAV RT, Rous Associated Virus (RAV) RT, and Myeloblastosis Associated Virus (MAV)
  • Patent Application 2003/0198944 (hereby incorporated by reference in its entirety). For review, see e.g., Levin, 1997, Cell, 88:5-8; Brosius et al.51995, Virus Genes 11:163-79.
  • Known reverse transcriptases from viruses require a primer to synthesize a DNA transcript from an RNA template.
  • Reverse transcriptase has been used primarily to transcribe RNA into cDNA, which can then be cloned into a vector for further manipulation or used in various amplification methods such as polymerase chain reaction (PCR), nucleic acid sequence-based amplification (NASBA), transcription mediated amplification (TMA), or self-sustained sequence replication (3SR).
  • PCR polymerase chain reaction
  • NASBA nucleic acid sequence-based amplification
  • TMA transcription mediated amplification
  • 3SR self-sustained sequence replication
  • sample generally refers to a biological sample of a subject.
  • the biological sample may comprise any number of macromolecules, for example, cellular macromolecules.
  • the sample may be a cell sample.
  • the sample may be a cell line or cell culture sample.
  • the sample can include one or more cells.
  • the sample can include one or more microbes.
  • the biological sample may be a nucleic acid sample or protein sample.
  • the biological sample may also be a carbohydrate sample or a lipid sample.
  • the biological sample may be derived from another sample.
  • the sample may be a tissue sample, such as a biopsy, core biopsy, needle aspirate, or fine needle aspirate.
  • the sample may be a fluid sample, such as a blood sample, urine sample, or saliva sample.
  • the sample may be a skin sample.
  • the sample may be a cheek swab.
  • the sample may be a plasma or serum sample.
  • the sample may be a cell-free or cell free sample.
  • a cell-free sample may include extracellular polynucleotides. Extracellular polynucleotides may be isolated from a bodily sample that may be selected from blood, plasma, serum, urine, saliva, mucosal excretions, sputum, stool and tears.
  • Sequequencing generally refers to methods and technologies for determining the sequence of nucleotide bases in one or more polynucleotides. Any method of sequencing known in the art may be used to evaluate the products of a reaction performed by an engineered reverse transcriptase of the current application. Sequencing can be performed by various systems currently available, such as, without limitation, a sequencing system by Illumina ® , Pacific Biosciences (PacBio ® ), Oxford Nanopore ® , or Life Technologies (Ion Torrent ® ).
  • sequencing may be performed using nucleic acid amplification, polymerase chain reaction (PCR) (e.g., digital PCR, quantitative PCR, or real time PCR), or isothermal amplification.
  • PCR polymerase chain reaction
  • a read may include a string of nucleic acid bases corresponding to a sequence of a nucleic acid molecule that has been sequenced.
  • systems and methods provided herein may be used with proteomic information.
  • substantially complementary means that a first sequence is at least 50%, 55%, 60%, 65%, 70%, 75%, 80%, 85%, 90%, 95%, 97%, 98% or 99% identical to the complement of a second sequence over a region of 2, 3, 4, 5, 6, 7, 8, 9, 10, 11, 12, 13, 14, 15, 16, 17, 18, 19, 20-40, 40-60, 60-100, or more nucleotides, or that the two sequences hybridize under stringent hybridization conditions.
  • Substantially complementary also means that a sequence in one strand is not completely and/or perfectly complementary to a sequence in an opposing strand, but that sufficient bonding occurs between bases on the two strands to form a stable hybrid complex in set of hybridization conditions (e.g., salt concentration and temperature). Such conditions can be predicted by using the sequences and standard mathematical calculations known to those skilled in the art.
  • the term “Subject,” as used herein, generally refers to an animal, such as a mammal (e.g., human) or avian (e.g., bird), or other organism, such as a plant.
  • the subject can be a vertebrate, a mammal, a rodent (e.g., a mouse), a primate, a simian or a human.
  • Animals may include, but are not limited to, farm animals, sport animals, and pets.
  • a subject can be a healthy or asymptomatic individual, an individual that has or is suspected of having a disease (e.g., cancer) or a pre-disposition to the disease, and/or an individual that is in need of therapy or suspected of needing therapy.
  • a subject can be a patient.
  • a subject can be a microorganism or microbe (e.g., bacteria, fungi, archaea, viruses).
  • Thermoreactivity refers to the ability of a reverse transcriptase to exhibit enzyme activity at elevated temperatures.
  • “Thermostability” or “thermostable” refers to the ability of a reverse transcriptase to withstand exposure to elevated temperatures, but not necessarily show activity at such elevated temperatures.
  • thermostable reverse transcriptase or polymerase refers to any enzyme that catalyzes polynucleotide synthesis by addition of nucleotide units to a nucleotide chain using DNA or RNA as a template and has an optimal activity at a temperature above 53° C.
  • unique molecular identifier As used herein, the terms “Unique molecular identifier”, “Unique molecular identifying sequence”, “UMI” and “UMI sequence” are used synonymously.
  • Individual barcoded molecules may comprise a common barcode sequence such as a partition specific sequence or a spatial array where every capture probe has a unique barcode sequence.
  • flanking sequence is intended a nucleic acid sequence capable of binding to an analyte.
  • Variant means a protein which is derived from a precursor protein (such as the native protein, for example MMLV native protein as set forth in SEQ ID NO:7) by addition of one or more amino acids to either or both the C- and N-terminal end, substitution of one or more amino acids at one or a number of different sites in the amino acid sequence, deletion of one or more amino acids at either or both ends of the protein or at one or more sites in the amino acid sequence, or addition of a fusion domain.
  • SEQ ID NO:1 is a variant of MMLV.
  • an enzyme variant is preferably achieved by modifying a DNA sequence which encodes for the wild-type protein, transformation of that DNA sequence into a suitable host, and expression of the modified DNA sequence to form the derivative enzyme.
  • a variant reverse transcriptase of the invention includes a protein comprising altered amino acid sequences in comparison with a precursor enzyme amino acid sequence wherein the variant reverse transcriptase retains the characteristic enzymatic nature of the precursor enzyme but which may have altered properties in some specific aspect.
  • an engineered reverse transcriptase variant may have an altered pH optimum or increased temperature stability but may retain its characteristic transcriptase activity.
  • a “Variant” may have at least about 45%, at least about 50%, at least about 55%, at least about 60%, at least about 65%, at least about 70%, at least about 75%, at least about 80%, at least about 85%, at least about 88%, at least about 90%, at least about 91 %, at least about 92%, at least about 93%, at least about 94%, at least about 95%, at least about 96%, at least about 97%, at least about 98%, at least about 99%, or at least about 99.5% sequence identity to a polypeptide sequence when optimally aligned for comparison.
  • Percent identity may pertain to the percent identity of the DNA binding domain or the engineered reverse transcriptase portion of an engineered reverse transcriptase.
  • a variant residue position is described in relation to the wild-type or precursor amino acid sequence set forth in SEQ ID NO:7; the amino acid position is indexed to SEQ ID NO:7.
  • a fusion variant comprises at least one fusion domain selected from DNA binding domains described elsewhere herein.
  • a protein having a certain percent (e.g., at least 90%, at least 91%, at least 92%, at least 93%, at least 94%, at least 95%, at least 96%, at least 97%, at least 98%, or at least 99%) of sequence identity with another sequence means that, when aligned, that percentage of bases or amino acid residues are the same in comparing the two sequences.
  • This alignment and the percent homology or identity can be determined using any suitable software program known in the art, for example those described in CURRENT PROTOCOLS IN MOLECULAR BIOLOGY, Ausubel et al., eds., 1987, Supplement 30, section 7.7.18.
  • Representative programs include the Vector NTI AdvanceTM 9.0 (Invitrogen Corp. Carlsbad, CA), GCG Pileup, FASTA (Pearson et al. (1988) Proc. Natl Acad. ScL USA 85:2444-2448), and BLAST (BLAST Manual, Altschul et al., Nat’l Cent. Biotechnol. Inf., Nat’l Lib. Med. (NCIB NLM NIH), Bethesda, Md., and Altschul et al., (1997) Nucleic Acids Res.25:3389-3402) programs.
  • Another typical alignment program is ALIGN Plus (Scientific and Educational Software, PA), generally using default parameters.
  • sequence alignment software programs that find use are the TFASTA Data Searching Program available in the Sequence Software Package Version 6.0 (Genetics Computer Group, University of Wisconsin, Madison, WI and CLC Main Workbench (Qiagen) Version 20.0. The present disclosure is not limited to the software being used to align two or more sequences.
  • WT Wild-type or “WT” refers to a gene or gene product that has the characteristics of that gene or gene product when isolated from a naturally occurring 143 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC source.
  • amino acid sequence set forth in SEQ ID NO:7 is a WT Murine Moloney Leukemia Virus (MMLV) sequence (Genbank NP_955591.1 p80 RT).
  • MMLV Murine Moloney Leukemia Virus
  • Genbank NP_955591.1 p80 RT WT Murine Moloney Leukemia Virus
  • nucleic acids are written left to right in 5' to 3' orientation; amino acid sequences are written left to right in amino to carboxy orientation, respectively.
  • the headings provided herein are not limitations of the various aspects or embodiments of the invention which can be had by reference to the specification as a whole. Accordingly, the terms defined immediately below are more fully defined by reference to the specification as a whole. EXAMPLES [000525] It will be understood that the reference to the below examples is for illustration purposes only and do not limit the scope of the claims.
  • Exemplary engineered RTs comprising a MMLV RT variant (42BL) operably linked N-terminally to a DNA binding protein derived from a budding yeast DAT1 were generated.
  • engineered RT comprising C-terminal fusions can also be constructed. Dat from different budding yeasts that contains at least three repeated pentads of G-R-K-P-G can be used.
  • DAT1 N-terminal fusion protein was generated; and an DAT1C-terminal fusion protein was generated.
  • the DAT1 fusion proteins are produced with an N-terminal 6x His Tag and thrombin cleavage site.
  • the 6x His Tag was used for purification purposes and removed by thrombin cleavage.
  • Exemplary engineered RTs e.g., variant MMLV
  • RT operably linked N-terminally to a truncated DAT1 molecule comprising 90 amino acid was also generated. While the majority of exemplary engineered RT tested below are of N-terminal fusions, C-terminal RT fusions can also be constructed.
  • homologs of DAT1 from other organisms can also be used to generate the engineered reverse transcriptase described herein as shown by N-DAT1-TL-QID04042BL; N-DAT1-TL-XP36142BL; N-DAT1-TL- XP55842BL; or N-DAT1-TL-XP68342BL.
  • DAT1(90) can be fused to other reverse transcriptase enzymes, such as other variants of MMLV reverse transcriptases as shown by N-DAT-9042B; N-DAT-9050A+ G; N-DAT-90 SOLD 33 VDG; N-DAT-90 SOLD 01; C-DAT-90 SOLD 01.
  • Example 1 Capillary Electrophoresis Assay Validation [000531] Reverse transcription and sequencing reactions were prepared. The reaction volume was 50 ⁇ l and reactions contained 5’-end labeled FAM Reverse Transcriptase primer 2, RT Reagent B (Chromium Next GEM Single Cell Reagent, 10X Genomics), RNA template (RNA Temp 2), template switching oligo 1 (TSO1), and the indicated engineered reverse transcriptase.
  • Table 3A Capillary Electrophoresis Assay Reactants R k Fi l 145 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC
  • Table 3A Capillary Electrophoresis Assay Reactants Reagent Stock Final on 5’ kit (10X Genomics, Inc), except the reverse transcriptase was altered for a particular reaction. Stock concentrations and final concentrations in the reactions are shown in Tables 3A-B. Variations of the assay stock concentrations and final concentrations in the reactions shown in Table 4 were used. The reactions included stoichiometrically equal amounts of enzyme and template for single turnover conditions.
  • Samples were loaded on a SeqstudioTM (Thermo Fisher Scientific) and fragment analysis by capillary electrophoresis was carried out with the appropriate dye channels and size standards.
  • the assay was validated with synthetically sized oligonucleotides and with a transcription positive, template switching null engineered reverse transcriptase and a transcription positive, template switching positive reverse transcriptase (Enzyme Mix C,).
  • the GEM-U reagent approximates the formulation of the actual reagent mixture in a GEM assay 146 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC when the contents of the Z1 and Z2 channels are mixed.
  • Capillary Electrophoresis Assay Reactants are disclosed in Table 1A, Capillary Electrophoresis Assay template, Primer and TSO sequences are shown in Table 4A.
  • Reverse transcription and sequencing reactions were also prepared using GAPDH or GRCh38 as a template. The reaction volume was 50 ⁇ l; reactions contained 5’-end labeled GAPDH or GRCh38 primer, GEM-U reagent, RNA template (GAPDH or GRCh38 template), template switching oligo 1 (TSO1), and the indicated engineered reverse transcriptase. Stock concentrations and final concentrations in the reactions are shown in Table 4B. The reactions included stoichiometrically equal amounts of enzyme and template for single turnover conditions.
  • reaction volume was 50 ⁇ l; reactions contained 5’-end labeled GAPDH primer, GEM-U reagent, RNA template (GAPDH template), template switching oligo 1 (TSO1), and the indicated engineered reverse transcriptase(s).
  • the final concentrations in the reactions are shown in Table 4B.
  • the reaction buffer was SOP for SC-5’ and the reaction time was 45 minutes.
  • Tables 4A-B show Capillary Electrophoresis (CE) Assay Reactants and Template, Primer and TSO sequences (SEQ ID NOS:173, 175, 176, respectively in order of appearance.)
  • Table 4A Capillary Electrophoresis Assay Template, Primer and TSO sequences 147 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC
  • Table 4A Capillary Electrophoresis Assay Template, Primer and TSO sequences Reagent Stock Final Volume [000539] Several mutants were constructed using a Q5 mutagenesis kit (NEB) with mutagenic primers per manufacturing instructions.
  • NEB Q5 mutagenesis kit
  • Amplification conditions were an initial denaturation at 95°C for 2.5 minutes, 30 cycles of denature (95°C, 30 sec), a 45 sec gradient annealing and extension at 72°C for 6 minutes, 35 sec, followed by a final extension at 72°C for 2 minutes.
  • Amplification reactions with multiple annealing gradient temperatures (65.2°C, 67°C, 68.5°C and 69.6°C) were performed.
  • Amplification products were evaluated on a 1.2% agarose E-Gel using SYBR-Safe. Products were pooled prior to clean-up. Cloning and expression were performed in the Acella cell line from EdgeBio (San Jose, CA). Cells were selected on LB-Kanamycin plates.
  • DAT1 N- terminal and C-terminal fusions to an engineered reverse transcriptase having the amino acid sequence set forth in SEQ ID NO:1 were obtained by screening of bacterial colonies. The sequences of the fusion proteins were confirmed using method known in the art. 148 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC
  • Example 3 Single Cell Sensitivity and Mapping [000542]
  • PBMCs peripheral blood monocytes
  • C57B/L6 mouse peripheral blood monocyte cells
  • Results from engineered reverse transcriptase were compared to results obtained from a commercially available engineered MMLV or MMLV variants (SOLD 001 (SEQ ID NO: 65); and SOLD 33 VDG (SEQ ID NO: 173).
  • SOLD 001 SEQ ID NO: 65
  • SOLD 33 VDG SEQ ID NO: 173
  • FIG.20 and FIG.21 show that DAT1 in combination with reverse transcriptase (MMLV, 42B, or other 42B variant thereof (e.g., SOLD 001 or SOLD 033 VDG)), either fused at the N-terminus or C-terminus of the RT improved GEX sensitivity even at low sequencing depth. Improvement was observed with both gene expression, which was increased by up to ⁇ 37% and UMI , which was increased by up to 13% as captured at 20k rrpc. Gains were even more significant at higher read depth.
  • MMLV, 42B, or other 42B variant thereof e.g., SOLD 001 or SOLD 033 VDG
  • the engineered RT disclosed herein exhibited large change in differential gene expression in single cell assays.
  • the engineered RT molecules comprising DAT1 or variant thereof disclosed herein picked-up to about 5000 additional genes when compared to a non-DAT1 RT (e.g., 42B).
  • FIGs.20-26 These engineered RT polypeptide also exhibited increase in median UMI counts per spot and median gene counts per spot in spatial assay.
  • the engineered RT molecules gave decrease in fraction of reads mapped to exons with gain in fraction mapped to introns. See e.g., FIG.21. A performance difference between FPLC and plate purified proteins was performed.
  • RT enzymes tested included 42B (SEQ ID NO: 1, SEQ ID NO: 143, or SEQ ID NO: 172), 50A+G (Table 2; SEQ ID NO: 147), 42B_L (Table 2; SEQ ID NO: 145).
  • FIGs.22A-B show the relative differences in performance of the engineered RT polypeptides or control non-DAT RT compared to control RT (42B).
  • the median genes and UMIs/cell at 50k raw-reads per cell were used to compare the sensitivity of three reverse transcriptases with and without the DAT fusion domain.42B-NDAT showed 43.29% (Median genes/cell) and 39.80% (median UMIs/cell) enhancement over 42B alone.
  • 42BL-CDAT showed 30.07% (Median genes/cell) and 23.76% (median UMIs/cell) enhancement over 42B alone.
  • 42BL-NDAT showed 47.40% (Median genes/cell) and 41.32% (median UMIs/cell) enhancement over 42B alone.
  • 50A+G-NDAT showed 45.00% (Median genes/cell) and 40.63% (median UMIs/cell) enhancement over 42B alone.
  • non-DAT RT, 42 B L only showed 7.25% (Median genes/cell) and 13.91% (median UMIs/cell) enhancement over 42B alone.
  • the non-DAT RT, 50A+G only showed 28.66% (Median genes/cell) and 38.01% (median UMIs/cell) enhancement over 42B alone.
  • both a C-terminal fusion and an N-terminal fusion of DAT to 42BL increased median genes per cell and median UMIs per cell in the single cell assays.
  • the C- terminal fusion increased median genes per cell by 21% compared to 42BL (3032 versus 2500) and median UMIs per cell by 9 % (8976 versus 8262) as compared to 42BL.
  • the N-terminal fusion increased median genes per cell by 37% (3436 versus 2500) and median UMIs per cell by 24%.
  • the 50A+G-DAT fusion showed a 13 % increase in genes per cell (3380 versus 2999) and a 2% increase in UMIs per cell (10200 versus 10010).
  • each DAT RT 150 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC fusion tested demonstrated improved quality metrics in a transcriptomics assay as compared to a matched control including the same RT but lacking the DAT domain.
  • These results showed clear performance gains in engineered RT polypeptides comprising a DAT fusion on either the N-terminal or C-terminal domain.
  • results also demonstrated the enhanced sensitivity of a single cell assay using an engineered reverse transcriptase polypeptide or an engineered recombinant reverse transcriptase comprising DAT1 described herein as shown by the median genes identified per cell (Median genes/cell) or the median UMIs identified per cell (median UMIs/cell).
  • an engineered RT of the present disclosure significantly improved the RT sensitivity when compared to a non-DAT1 RT; and an engineered reverse transcriptase comprising a DAT1 binding domain further significantly increased the gain in sensitivity of the engineered reverse transcriptase described herein.
  • FIG.23A The performance of various engineered RT disclosed herein were also analyzed at maximum normalization depth using the median genes/cell (FIG.23A) or the median UMIs/Cell (FIG.23B) to compare the generated library complexity of three reverse transcriptases (42B, 42B L, and 50A+G) with (N-DAT or C-DAT) and without the DAT domain.
  • FIGs.23 C-D show saturation curves of the median genes (FIG.23C) and counts/cell (FIG.27C) as a function of read depth, which further demonstrate that the median genes and counts/cell were higher using the engineered RT with a DAT DNA binding domain when compared to MMLV variants lacking the DAT DNA binding domain.
  • FIG.23 demonstrates a clear benefit of using the DAT DNA binding domain in a Single Cell 5’ (SC-5’) gene expression assay.
  • SC-5 Single Cell 5’
  • FIGs.24A-F and FIGs.25A-F showed significant levels of differential gene expression with 42B L-CDAT, 42B L-NDAT, 50A+G-NDAT.
  • FIGs.24A-F shows the differential gene expression of some engineered RT comprising DAT1 at the N- terminus.
  • FIGs.24A, C, and E feature scatter plots showing gene expression correlation of three reverse transcriptases with and without the DAT fusion domain.
  • FIGs.24B, D, and F feature 151 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC volcano plots showing the number of differentially expressed genes between three reverse transcriptases with and without the DAT fusion domain.
  • FIGs.25A-F illustrate the differential gene expression of some engineered RT comprising DAT1 at the C-terminus.
  • FIGs.25A, C, and E feature scatter plots showing gene expression correlation comparison of engineered RT comprising N- Terminal DAT Domain and C- Terminal DAT Domain.
  • FIGs.25B, D, and F feature volcano plots comparing the number of differentially expressed genes between various RT. Gains in performance, were still present, but not as significant with a C-terminal fusion.
  • the C-DAT terminal fusion did not perform as well as the N-terminal DAT fusion, but both N-terminal DAT fusion and C-DAT terminal fusion showed performance gains when compared to the RT backbone (42B L; SEQ ID NO: 145) alone.
  • the conditions become more similar in complexity (42B L - NDAT > 42B L - CDAT > 42B L, see FIGs 22-23), there are lower levels of DEGs and the feature scatter plots begin to correlate better.
  • FIGs.26A-D show graphs illustrating the performance comparison of the impact of the DAT fusion domain across three reverse transcriptase backbones based on median genes (FIG.26A) and UMIs/cell at maximum normalization depth (FIG.26B), gene expression correlation (FIG.26C), and differential gene expression (FIG.26D).
  • the aggregated metrics comparing the performance among 42B, 42B L, and 50A+G backbones with and without the DAT fusion showed a clear performance benefit from the DAT domain, such as e.g., enhanced sensitivity. Therefore, the 42B L-DAT1 fusion RT significantly improve in single cell assay performance.
  • Example 5 Assays for analyses of Engineered RT polypeptides
  • Any of the engineered RT enzymes of the invention including without limitation any of the enzymes described in Table 1, or Table 2 could be analyzed in any suitable assay, including without limitation the assays described herein.
  • Assays include without limitation 5’ gene expression analyses, with or without VDJ analysis, 3’ gene expression analysis, epigenetic analysis, or multiomic analyses.
  • experiments are carried out as found in the manufacturer’s instructions for the Chromium Single Cell 5’ Gene Expression 152 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC Assay kit (10X Genomics); Chromium Single Cell 3’ Gene Expression Assay kit (10X Genomics), including any of multiomic extensions or applications.
  • Example 6 Single Cell 3’ and 5’ cDNA Yields [000556]
  • PBMCs peripheral blood monocytes
  • Emulsion droplets can contain gel beads with either barcoded poly-dT primer sequences (3’ configuration) or barcoded template switch oligo sequences (5’ configuration) that also include a UMI and Illumina ® Read 1 sequence.
  • the poly-dT primer hybridizes to the poly-A tail of the cellular mRNA, which is extended by the reverse transcriptase. Once the end of the template is reached, the reverse transcriptase will exhibit terminal transferase activity to add an overhang of three non-templated deoxycytidines (CCC) to the 3’ end of the synthesized cDNA.
  • CCC non-templated deoxycytidines
  • the CCC overhang will hybridize to the 3 riboguanosines (rGrGrG) present on the 3’ end of the template switch oligo, allowing the reverse transcriptase to “switch” templates and continue synthesis to the 5’ end of the template switch oligo.
  • the barcode and UMI will allow either the 3’ or 5’-end of the mRNA molecule to be identified in the final sequencing library.
  • cDNA was then amplified via PCR, purified with a 0.6x SPRI, and quantified with an Agilent Bioanalyzer using the DNA High Sensitivity Kit. The cDNA yield (ng) was then obtained.
  • PBMCs peripheral blood monocytes
  • Either 10 ⁇ L of the amplified cDNA (3’ conditions) or 20 ⁇ L containing a maximum of 50 ng of amplified cDNA (5’ conditions) can then be fragmented and A-tailed, cleaned with a double-sided SPRI (0.6x/0.8x), ligated to functional adaptors with an Illumina ® Read 2 sequence, cleaned with a 0.8x SPRI, and then can be further amplified with sample indexing primers that include the P5 and P7 priming sites and the i5 and i7 sample indexes.
  • the amplification product can be cleaned up with a double-side (0.6x/0.8x) SPRI, and the average size can be determined with an Agilent Bioanalyzer using the DNA High Sensitivity Kit.
  • the material can then be quantified by qPCR 153 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC and pooled for next generation sequencing on an Illumina ® Novaseq targeting a sequencing depth of at least 50,000 reads per cell and using the following run parameters (Read 1: 28 cycles, i7 Index: 10 cycles, i5 Index: 10 cycles, Read 2: 90 cycles). Data can be collected, demultiplexed, and processed. Standard quality metrics were obtained. [000558] Generally, the single cell 5’ reactions use less enzyme and TSO oligo than the single cell 3’ reactions.
  • the 5’ TSO oligo is also twice the length of the 3’ TSO oligo with varied sequence context due to the presence of the UMI and the barcode.
  • the single cell 5’ reaction conditions are generally considered a more stringent test of performance than the 3’ single cell reaction conditions.
  • FIG.21 summarizes results from a series of experiments using the 5’ reaction conditions. The figure summarizes metrics of the 5’ single cell experiments, including 20k read metrics, 50K read metrics and reads mapped to the transcriptome.
  • the engineered reverse transcriptase variants have the amino acid sequences provided set forth in SEQ ID NO: 65 (SOLD 001), SEQ ID NO: 173 (SOLD 33 VDG), SEQ ID NO: 174 (C-DAT 42BL), and SEQ ID NO: 175 (N-DAT 42BL).
  • the percent indicates the percent change from the results obtained with an engineered reverse transcriptase having the amino acid sequence set forth in SEQ ID NO:1 or 179 (42B).
  • the engineered reverse transcriptase polypeptides showed a significant improvement in sensitivity.
  • FIG.21 shows additional metrics related to results obtained from the indicated engineered reverse transcriptase in single cell 5’ experiments.
  • Immune profiling is an extension of the 5’ chemistry to profile genes specifically for T-cell and/or B-cell receptors in the mRNA pool.
  • Immune profiling are known in the art and generally include additional rounds of PCR on the cDNA with a pool of sequence specific primers to allow for targeted enrichment of T-cell and/or B-cell receptor genes. Immune profiling assays may also detect UMIs for B-cell receptor genes, namely IGH, IGK, and IGL (Immunoglobulin heavy chain (IGH), kappa (IGK), and light (IGL) chain). Immune profiling data is informative for immunology research and is an extension of standard gene expression evaluation. Methods of immune profiling include but are not limited to Chromium Next Gen Single CellTM kits (10X Genomics, Pleasanton CA).
  • Amplified cDNA (2 ⁇ l) from the 5’ configuration of reverse transcription reactions can be subjected to two additional rounds of PCR enrichment with TCR immune profiling, which included a double-sided (0.5x/0.8x) SPRI clean-up between the first and second round of thermal cycling reactions.
  • the amplified products can then be cleaned-up with a subsequent double-sided (0.5x/0.8x) SPRI, fragmented and A-tailed, ligated to functional adaptors with an Illumina ® Read 2 sequence, cleaned up with a 0.8x SPRI, and then further amplified with sample indexing primers that include the P5 and P7 priming sites and the i5 and i7 sample indexes.
  • the amplification product can be cleaned up with a 0.8x SPRI, and average size can be determined with an Agilent Bioanalyzer using the DNA High Sensitivity Kit.
  • the material can then be quantified by qPCR and can be pooled for next generation sequencing on an Illumina ® Novaseq targeting a sequencing depth of at least 5,000 reads per cell and using the following run parameters (Read 1: 28 cycles, i7 Index: 10 cycles, i5 Index: 10 cycles, Read 2: 90 cycles).
  • Data can be collected, demultiplexed, and single-cell V(D)J analysis can be performed. Results that are obtained from engineered reverse transcriptases can be compared to results are obtained from a commercially available enzyme or an RT lacking the DAT 1 DNA binding domain.
  • the percent change in median TRA UMI’s and median TRB UMI’s from mouse and human PBMCs for each RT tested can be shown as a percent change in median IGH, IGK and IGL from mouse PBMC’s. It is expected that the median TRA UMIs and median TRB UMIs obtained with any 155 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC engineered reverse transcriptase described herein h will be greater than those obtained with a commercially available reverse transcriptase or non-DAT RT in both human PBMCs and mouse PBMCs.
  • Example 9 Spatial assays [000565] This example shows various methods for performing spatial analysis using the engineered RT enzymes of the present disclosure.
  • the Example provides an exemplary method for detecting an individual gene expression or globally-expressed RNA transcripts in a fresh or frozen tissue using an engineered RT of SEQ ID NO: 175, which includes N-Terminal fusion of the first 90a.a.
  • N-DAT Dat 1
  • a 42B variant of SEQ ID NO: 1, 143, 145, or 172 compared to a control RT variant (having the same RT enzyme but lacking the N-DAT or SEQ ID NO: _1, 143, 145, or 172) using three different reverse transcription buffer formulations, the commercially available RT reagent, buffer X.4 and buffer X.5.
  • the commercially available RT reagent comprises 2% Glycerol, 50 mM Tris pH 8.3, 3 mM MgCl2, 75.04 mM KCl, 0.5% Supersonic F-108, 0.489 mg/ml BSA, and 3.16 mM dNTPs.
  • Buffer X.4 and buffer X.5 have the same components as the commercial RT reagent with the following differences.
  • Buffer X.4 contains 1.05% Glycerol, 158.3 mM NaCl, 0.579 mg/ml BSA, 4.685 mM dNTPs, 1.5mM dCTP, and 0.775mM GTP.
  • Buffer X.5 contains 133.8 mM NaCl, 1% Supersonic F-108, 3.96 mM dNTPs, 1.31 mM dCTP, and 0.655 mM GTP. [000566] Sample preparation.
  • a formalin-fixed, paraffin-embedded (FFPE) human tonsil sample, a frozen Human tonsil sample, or a fresh Human tonsil sample can be used for the method described herein.
  • Human tonsil FFPE sections on standard slides (for sandwich conditions; FIGs.1A-B, FIGs.2A-B, and FIGs.3A-C) or gene expression (GEx) slides (for non-sandwich control conditions) were deparaffinized, H&E stained, and imaged. After the imaging, the human tonsil sections were hematoxylin-destained with HCL solution. The sections were then decrosslinked by incubating at 70°C for 1 hour in decrosslinking solution.
  • the permeabilization solution was washed out and the human tonsil sample was prepared for analyte capture by adding 0.1X SSC buffer and subjected to a pre-equilibration thermocycling protocol (e.g., lid temperature and pre-equilibrate at 53 °C, reverse transcription at 53 °C for 45 minutes, and then hold at 4 °C).
  • a pre-equilibration thermocycling protocol e.g., lid temperature and pre-equilibrate at 53 °C, reverse transcription at 53 °C for 45 minutes, and then hold at 4 °C.
  • the SSC buffer was removed.
  • a Master Mix comprising a commercially available RT reagent buffer, buffer X.4, or buffer X.5, nuclease-free water, a template switch oligo, a reducing agent, and a reverse transcriptase (control 42B RT, N-DAT-42B RT, or small scale-purified N-DAT- 42B RT) was added to the human tonsil sample and subjected to a thermocycling protocol, which comprised performing a reverse transcription reaction at 53 °C for 45 minutes and hold at 4 °C.
  • a second strand synthesis was performed on the sample by subjecting the sample to a thermocycling protocol, which comprised e.g., pre-equilibrating at 65 °C, synthesizing the second strand at 65 °C for 15 minutes, then holding at 4 °C.
  • the Master Mix reagents were removed from the sample and 0.8M KOH was added and incubated for 5 minutes at room temperature. The KOH was removed and followed by the addition of an elution buffer.
  • a Second Strand Mix including a second strand reagent, a second strand primer, and a second strand-polymerizing enzyme, can be added to the sample and the sample can be sealed and incubated.
  • the reagents can be removed and elution buffer can be added and removed from the sample, and 0.8 M KOH can be added again to the sample and the sample can be incubated for 10 minutes at room temperature. Tris-HCl can be added and the reagents can be mixed. The sample can be transferred to a new tube, vortexed, and placed on ice.
  • cDNA amplification and quality control A qPCR Mix, including nuclease-free water, qPCR Master Mix, and cDNA primers, was prepared and pipetted into wells in a qPCR plate.
  • the sample was incubated 157 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC and thermocycled (e.g., lid temperature at 105 °C for -45-60 minutes; step 1: 98 °C for 3 minutes, step 2: 98 °C for 15 seconds, step 3: 63 °C for 20 seconds, step 4: 72 °C for one minute, step 5: [the number of cycles determined by qPCR Cq Values], step 6: 72 °C for 1 minute, and step 7: hold at 4 °C).
  • the sample can then be stored at 4 °C for up to 72 hours or at -20 °C for up to 1 week or use immediately.
  • a Fragmentation Mix including a fragmentation buffer and a fragmentation enzyme was prepared on ice. Elution buffer and fragmentation mix was added to each sample, mixed, and centrifuged. The sample mix was then placed in a thermocycler and cycled according to a predetermined protocol (e.g., lid temperature at 65 °C for ⁇ 35 minutes, pre-cool block down to 4 °C before fragmentation at 32 °C for 5 minutes, End-repair and A-tailing at 65 °C for 30 minutes and holding at 4 °C).
  • a predetermined protocol e.g., lid temperature at 65 °C for ⁇ 35 minutes, pre-cool block down to 4 °C before fragmentation at 32 °C for 5 minutes, End-repair and A-tailing at 65 °C for 30 minutes and holding at 4 °C.
  • SPRI Cleanup The 0.6X SPRI select Reagent was added to the sample and incubated for 5 minutes at room temperature.
  • the sample was placed on a magnet (e.g., in the high position) until the solution cleared, and the supernatant was transferred to a new tube strip.
  • 0.8X 158 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC SPRI select Reagent was added to the sample, mixed, and incubated for 5 minutes at room temperature.
  • the sample was placed on a magnet (e.g., in the high position) until the solution cleared.
  • the supernatant was removed and 80% ethanol was added to the pellet, incubated for 30 seconds, and washed. The ethanol wash was repeated and the sample placed on a magnet (e.g., in the low position) until the solution cleared.
  • elution buffer was added to the sample, mixed, and incubated for 2 minutes at room temperature.
  • the sample was placed on a magnet (e.g., in the high position) until the solution cleared, and a portion of the sample was moved to a new tube strip.
  • Ligation An Adaptor Ligation Mix, including ligation buffer, DNA ligase, and adaptor oligos, was prepared and centrifuged. The Adaptor Ligation Mix was added to the sample, pipette-mixed, and centrifuged briefly.
  • the average fragment size was determined using a Bioanalyzer trace or an Agilent TapeStation.
  • the library was sequenced using a sequencing platform, for example, MiSeq, NextSeq 500/550, HiSeq 2500, HiSeq 3000/4000, NovaSeq, and iSeq. See, Illumina ® , Indexed Sequencing Overview Guides, February 2018, Document 15057455v04; and Illumina ® Adapter Sequences, May 2019, Document #1000000002694vl 1, each of which is hereby incorporated by reference, for information on P5, P7, i7, i5, TruSeqTM Read 2, indexed sequencing, and other reagents described herein.
  • N-DAT RT variant improved the sensitivity of the spatial assay while maintaining good spatial resolution in human samples.
  • the effect of the 42B RT control variant (SEQ ID NO: 1, 143, 145, or 172) and the N- DAT-42B RT variant were compared using the three buffer types as previously described.
  • the human tonsil samples were permeabilized using a Kryptonite permeabilization buffer comprising 6% Sarkosyl, 1M urea, and 4ug/ul proteinase K.
  • the reverse transcriptase reaction was performed on human tonsil samples were prepared as disclosed above using six different conditions: (1) the 42B RT control variant was used with RT reagent; (2) the 42B RT control variant was used with Buffer X.5; (3) the engineered N-DAT-42B RT variant was used with the RT reagent; (4) the engineered N-DAT-42B RT variant was used with buffer X.4; (5) the engineered N-DAT-42B RT variant was used with buffer X.5; and (6) the small-scale purified engineered N-DAT-42B RT variant was used with buffer X.5. Most of the engineered N-DAT- 42B RT variants tested were purified by FPLC.
  • N-DAT-42B RT variants underwent a small scale purification. RT was performed using a 2 step protocol: 45 min at 53 °C and 30 min at 42 °C. [000580] As shown in FIGs.28A-F, the engineered N-DAT-42B RT variant maintained a good spatial resolution using all three buffers tested.
  • FIGs.28A-F show UMI heat maps showing globally detected gene expression in a human tonsil tissue analyzed using N-DAT-42B RT variant (FIGs.28C-F) or 42B RT variant (FIGs.28A-B) and the three different buffer formulations.
  • the NDAT1 RT variant was either FPLC-purified or underwent a small scale purification (FIG.28F).
  • the global detection of gene expression is shown as a heat map.
  • the RT Reagent (FIG.28A and FIG.28C) showed better mapping metrics when compared to buffer X.4 (FIG.28D), and buffer X.5 (FIG.28B, FIG.28E, and FIG.28F).
  • the commercially available RT reagent comprised 2% Glycerol, 50 mM Tris pH 8.3, 3 mM MgCl2, 75.04 mM KCl, 0.5% Supersonic F-108, 0.489 mg/ml BSA, and 3.16 mM dNTPs.
  • Buffer X.4 and buffer X.5 have the same components as the commercial RT reagent with the following differences.
  • Buffer X.4 contains 1.05% Glycerol, 158.3 mM NaCl, 0.579 mg/ml BSA, 4.685 mM dNTPs, 1.5mM dCTP, and 0.775mM GTP.
  • Buffer X.5 contains 133.8 mM NaCl, 1% Supersonic F-108, 3.96 mM dNTPs, 1.31 mM dCTP, and 0.655 mM GTP.
  • a MMLV RT variant without a DAT fusion domain was used as a control (e.g., a RT variant comprising the amino acid sequence of SEQ ID 160 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC NO: 1, 143, 145, or 172).
  • the NDAT1 RT variant was either FPLC-purified or underwent a small-scale purification (FIG.28F).
  • FIGs.28A-F show that NDAT1 RT variant improved the sensitivity of the assay while maintaining good spatial resolution when compared to the control RT variant, with all three buffer conditions.
  • FIGs.29A-F show that the engineered N-DAT-42B RT also increased the quality and the sensitivity metrics of the spatial assays.
  • FIGs.29A-F show graphs comparing the quality, sensitivity, and detection of gene expression under the same six conditions shown in FIGs.28A- F.
  • fraction reads in spots under tissue (FIG.29A) was substantially the same among the 6 conditions tested.
  • the smallest standard deviation across the six conditions tested was obtained using RT reagent and with the engineered N-DAT-42B RT.
  • FIG.29D Analysis of median UMI counts per spot (30k mapped spot-reads per spot, mapped to GRch38 reference genome assembly) (FIG.29D) showed that median UMI counts per spot were significantly enhanced in all three conditions using an N-DAT RT variant when compared to the three conditions using the RT control variant. Unexpectedly, N-DAT RT in bufferX.4 (N- DAT_RTX4) showed less variability. In addition, the GRch38 median genes per spot (30k mapped spot-reads per spot) (FIG.29E) analysis showed that median genes per spot were significantly enhanced in all samples comprising an N-DAT RT variant when compared to the 42B RT samples regardless of the buffer used.
  • engineered NDAT1 RT variant improved the sensitivity of the spatial assay while maintaining good spatial resolution when compared to the 42B RT variant (FIGs.29D-E).
  • the RT Reagent appeared to show better mapping metrics when compared to Buffer X.4, and Buffer X.5.
  • RT Reagent B with N-DAT at various concentrations and timings can also be tested and will likely show similar results as the RT Reagent. [000582]
  • Assessment of individual gene expression also showed substantially similar results as globally detected gene expression.
  • FIGs.30A-F show UMI heat maps shown as a log10(UMI) 161 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC ranging from cold to hot (0, 0.5, 1.0, 1.5, and 2.0), which illustrate the gene expression of RGS3 in a human tonsil tissue analyzed using N-DAT RT variant (FIGs.30C-F) or 42BRT control variant (FIGs.30A-B) under the three different buffer formulations (RT reagent (FIG.30A and FIG.30C), Buffer X.4 (FIG.30D), and Buffer X.5 (FIG.30B, FIG.30E, and FIG.30F)).
  • N-DAT RT variant FIGs.30C-F
  • 42BRT control variant FIGs.30A-B
  • FIGs.31A-C show UMI heat maps shown as a log10(UMI) ranging from cold to hot (0, 0.5, 1.0, 1.5, 2.0, and 2.5) illustrating the gene expression of KRT5 in a human tonsil tissue analyzed using N-DAT RT variant (FIGs.31C-F) or RT control variant (FIGs.31A-B) using three different buffer formulations (a commercially available RT reagent (FIG.31A and FIG. 31C), Buffer X.4 (FIG.31D), and Buffer X.5 (FIG.31B, FIG.31E, and FIG.31F)).
  • a commercially available RT reagent FIG.31A and FIG. 31C
  • Buffer X.4 FIG.31D
  • Buffer X.5 FIG.31B, FIG.31E, and FIG.31F
  • Example 11 Spatial resolution in non-human/mouse tissue.
  • SD Visium standard definition
  • HD Visium high definition
  • Visium HD slides contain two 6.5 x 6.5 mm Capture Areas with a continuous lawn of oligonucleotides arrayed in ⁇ 11 million 2 x 2 ⁇ m barcoded squares without gaps, achieving single cell–scale spatial resolution.
  • the data are output at 2 ⁇ m, as well as multiple bin sizes.
  • the 8 x 8 ⁇ m bin is the recommended starting point for visualization and analysis. See e.g., 10xgenomics.com/platforms/visium; 10xgenomics.com/products/visium-hd-spatial-gene-expression.
  • the zebrafish samples were permeabilized using a Kryptonite permeabilization buffer comprising 6% Sarkosyl, 1M urea, and 4ug/ul proteinase K.
  • the reverse transcriptase reaction was performed on zebrafish samples using T reagent, the engineered N-DAT-42B RT (NDAT- 42BL) variant, TSO, and dTT.
  • the RT was performed using a 2 step protocol-45 min at 53 °C and 30 min at 42 °C.
  • the single strand synthesis was performed using a control RT (e.g., MMLV TR variant comprising an amino acid sequence of SEQ ID NO: 1, 143, 145, or 172), RT Reagent, and SS Primers.
  • N-DAT fusion did not sufficiently process the synthesis of the second strand cDNA in the Zebrafish tissue. This was 162 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC likely caused by the N-DAT fusion configuration or the RT variant configuration. For example, the configuration may have slowed the enzyme from initiating or continuing the synthesis of the second strand. However, the N-DAT fusion sufficiently process the synthesis of the first strand cDNA in the Zebrafish tissue.
  • FIGs.32A-B show UMI heat maps shown as a count log10 ranging from cold to hot (0.5, 1, 1.5, 2, 2.5, 3, 4, 4.5, and 5), illustrating globally detected gene expression in Zebrafish (non-human or mouse tissue) using a spatial assay with a first slide configuration.
  • FIGs.32C-D show heat maps illustrating individual gene expression.
  • FIG.32C shows Crgm1 gene expression as log normalized per experiment ranging from cold to hot (0.00-5.0).
  • FIG.32D shows KRT5 gene expression as log normalized per experiment ranging from cold to hot (0.0-7.0).
  • FIG.32E shows Rho gene expression as log normalized per experiment ranging from cold to hot (0.0-8.0).
  • Fraction reads mapped to genome are the fraction of reads that mapped to a unique gene in the genome. In general, the read must be consistent with annotated splice junctions and are considered for UMI counting.
  • FIGs.33A-D also show UMI heat maps showing globally detected gene expression (FIG.33A) or individual gene expression (FIGs.33B-C) in Zebrafish (non-human or mouse tissue) analyzed using NDAT1 RT variant (42BL-N-DAT1) and RT reagent buffer under sandwich configuration conditions for a spatial assay with a second slide configuration.
  • FIG. 33B shows Crgm1 gene expression as log normalized per experiment ranging from cold to hot (0.00-4.0).
  • FIG.33C shows KRT5 gene expression as log normalized per experiment ranging from cold to hot (0.0-9.0).
  • FIG.33D shows Rho gene expression as log normalized per experiment ranging from cold to hot (0.0-7.0).
  • the quality, and sensitivity metrics were: fraction reads mapped to genome (0.72).
  • the engineered N-DAT RT variants disclosed herein enhanced the quality and the sensitivity of the spatial assay metrics while maintaining good spatial resolution when compared to a control RT variant.
  • the control RT variant lacked the N-DAT but comprised the same RT backbone as the engineered N-DAT RT variant.
  • the control RT variant consisted of the amino acid sequence of SEQ ID NO: 1, 142, 143, or 172.
  • Table 1 shows listing of non-limiting embodiments of RT enzymes of the present disclosure.
  • Tables 2 and 5 shows additional listing of amino acid and nucleic acid sequences of non-limiting embodiments of the engineered RTs of the present disclosure. Table 1.
  • Non-limiting embodiments of RT variants of the invention SEQ ID NO Description Sequences S L S P G V K H P K L T 165 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC Table 1.
  • Non-limiting embodiments of RT variants of the invention SEQ ID NO Description Sequences I K E G E L I A E N A 166 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC Table 1.
  • Non-limiting embodiments of RT variants of the invention SEQ ID NO Description Sequences aa ga tc tc cc g at c 167 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC Table 1.
  • Non-limiting embodiments of RT variants of the invention SEQ ID NO Description Sequences g c g c I V P GI I A T T P K W Q Y H P I V P GI I A T P K W Q Y H I V P GI I 168 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC Table 1.
  • Non-limiting embodiments of RT variants of the invention SEQ ID NO Description Sequences A T P K W Q Y H I P P G P F L P H S G P I P P P G P F L P H S G P 169 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC Table 1.
  • Non-limiting embodiments of RT variants of the invention SEQ ID NO Description Sequences I V P GI I A T P K T W Q Y H I V P GI I A T P K T W Q Y H V F S F K L 170 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC Table 1.
  • Non-limiting embodiments of RT variants of the invention SEQ ID NO Description Sequences E V F S F K L E M V F S F K L E N A V F S 171 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC Table 1.
  • Non-limiting embodiments of RT variants of the invention SEQ ID NO Description Sequences F K L E V F S D C Q G A G A V F S D C Q G A G K 172 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC Table 1.
  • Non-limiting embodiments of RT variants of the invention SEQ ID NO Description Sequences D V F S F K L E I V P GI I A T P K T W Q Y H I V P GI I A T P 173 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC Table 1.
  • Non-limiting embodiments of RT variants of the invention SEQ ID NO Description Sequences K T W Q Y H I P P G P F L P H S G P I P P P G P F L P H S G P I V P GI I A 174 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC Table 1.
  • Non-limiting embodiments of RT variants of the invention SEQ ID NO Description Sequences T P K W Q Y H I V P GI I A T P K T W Q Y H I V P GI I A T P K W Q Y H I V 175 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC Table 1.
  • Non-limiting embodiments of RT variants of the invention SEQ ID NO Description Sequences P GI I A T P K W Q Y H I V P GI I A T P K W Q Y H I P P G P F L P H S G 176 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC Table 1.
  • Non-limiting embodiments of RT variants of the invention SEQ ID NO Description Sequences P I P P G P F L P H S G P I V P GI I A T P K W Q Y G I V P G I A T P K 177 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC Table 1.
  • Non-limiting embodiments of RT variants of the invention SEQ ID NO Description Sequences W Q Y G I V P GI I A T P K W Q Y H I V P G I A T P K W Q Y H I V P GI I A T 178 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC Table 1.
  • Non-limiting embodiments of RT variants of the invention SEQ ID NO Description Sequences P K W Q Y H L L R L Q E F R L A K Q L L R L Q E F R L A K Q 179 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC Table 1.
  • Non-limiting embodiments of RT variants of the invention SEQ ID NO Description Sequences I P P G P F L P H S G P S P P P G P F L P H S G P K K P IS K P P 180 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC Table 1.
  • Non-limiting embodiments of RT variants of the invention SEQ ID NO Description Sequences T W Q Y H P 181 4876-6828-0003.1 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1 - 1: 0 . f 0 e 1: R.

Landscapes

  • Life Sciences & Earth Sciences (AREA)
  • Health & Medical Sciences (AREA)
  • Chemical & Material Sciences (AREA)
  • Organic Chemistry (AREA)
  • Zoology (AREA)
  • Genetics & Genomics (AREA)
  • Bioinformatics & Cheminformatics (AREA)
  • Wood Science & Technology (AREA)
  • Engineering & Computer Science (AREA)
  • General Health & Medical Sciences (AREA)
  • Biochemistry (AREA)
  • General Engineering & Computer Science (AREA)
  • Molecular Biology (AREA)
  • Medicinal Chemistry (AREA)
  • Biomedical Technology (AREA)
  • Biotechnology (AREA)
  • Microbiology (AREA)
  • Micro-Organisms Or Cultivation Processes Thereof (AREA)
  • Peptides Or Proteins (AREA)

Abstract

The disclosure provides recombinant reverse transcriptases (e.g, engineered reverse transcriptase polypeptides) comprising one or more DNA binding domains conjugated to a wild-type or a mutant reversed transcriptases that exhibit one or more altered reverse transcriptase related activities, such as but not limited to, altered template switching efficiency, altered transcription efficiency or both. The disclosure further provides compositions and kits comprising the recombinant reverse transcriptases (e.g., engineered reverse transcriptase polypeptides) and methods of producing, amplifying, or sequencing nucleic acid molecules using the recombinant reverse transcriptases

Description

Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    SEQUENCE SPECIFIC DNA BINDING PROTEINS CROSS REFERENCE [0001] This application claims priority from U.S. Provisional Patent Application No. 63/472,726, filed June 13, 2023, and U.S. Provisional Patent Application No.63/622,402, filed January 18, 2024. The entire contents of which are hereby incorporated by reference for all purposes. FIELD OF THE INVENTION [0002] The present invention relates to the field of protein engineering, particularly development of recombinant reverse transcriptase variants that exhibit one or more improved properties of interest. BACKGROUND [0003] A variety of single cell processing and analytical methods and systems are known in the art, including analysis of specific individual cells, analysis of different cell types within populations of differing cell types, analysis and characterization of large populations of cells for environmental, human health, or epidemiological forensic. However, these methods and systems remain inefficient for a variety of reasons, including spatial heterogeneity and insufficient transcript capture. [0004] Specifically, cells within a tissue can have different morphology and/or function due to varied analyte levels (e.g., gene and/or protein expression). The specific position of a cell within a tissue (e.g., the cell’s position relative to neighboring cells or the cell’s position relative to the tissue microenvironment) can also affect the cell’s morphology, differentiation, fate, viability, proliferation, behavior, signaling, and cross-talk with other cells in the tissue. [0005] This spatial heterogeneity has been previously studied using techniques that only provide data for a small handful of analytes in the context of an intact tissue or a portion of a tissue. Specifically, these techniques can provide substantial analyte data for dissociated tissues (i.e., single cells). However, they fail to provide information regarding the position of a single cell in a biological sample (e.g., tissue sample) caused in part by inefficient transcript capture. Accordingly, there is a need for improved single cell processing and analytical methods with 1 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    enhanced transcript capture. Specifically, there is a need for improved reverse transcriptases with improved sensitivity, efficiency, processivity, and transcript capture for single cell and/or spatial analysis applications. The present disclosure addresses this need. SUMMARY [0006] One aspect of the present disclosure provides an engineered reverse transcriptase (RT) polypeptide comprising: (a) an RT polypeptide sequence; (b) a DNA binding domain, where the DNA binding domain is from a molecule capable of binding a minor groove of a nucleic acid; and (c) a linker connecting the RT polypeptide sequence and the DNA binding domain. [0007] In some embodiments of the engineered RT polypeptide described herein, the DNA binding domain is located at the N-terminus of the RT polypeptide sequence. In some embodiments, the DNA binding domain is located at the C-terminus of the RT polypeptide sequence. In some embodiments, the linker is G(n)S(m)G(p), and (a) n = 0, 1, 2, 3, 4, 5,6 , 7, 8, 9, 1011, 12, 13, 14, 15, 16, 17, 18, 19 or 20; (b) m = 0, 1, 2, 3, 4, 5,6 , 7, 8, 9, 1011, 12, 13, 14, 15, 16, 17, 18, 19 or 20; (c) p = 0, 1, 2, 3, 4, 5, 6, 7, 8, 9, 1011, 12, 13, 14, 15, 16, 17, 18, 19 or 20; and (d) n, m, and p are selected independently. In some embodiments, the linker is a glycine-serine (GS) linker selected from the group consisting of (GS)n, (GSGGS)n, (SGGSG)n, (GGGS)n, (GGSG)n, (GGSGG)n, (GSGSG)n, (GSGGG)n, GGGSG)n, and (GSSSG)n, where n represents an integer of at least 1. In some embodiments, the linker is GGGS. In some embodiments, the linker is SGGSG. [0008] In some embodiments of the engineered RT polypeptide described herein, the DNA binding domain specifically recognizes adenine-thymine-rich region on a nucleic acid molecule. In some embodiments, the DNA binding domain specifically recognizes oligo(dA) or oligo(dT) tracts on a nucleic acid molecule. In some embodiments, the DNA binding domain comprises at least one AT-rich interaction domain. In some embodiments, the DNA binding domain comprises at least 2, at least 3, at least 4, at least 5, at least 6, at least 7, at least 8, at least 9, or at least 10 AT-rich interaction domains. [0009] In some embodiments, the AT-rich interaction domain comprises a core sequence, wherein the core sequence is a two base core sequence, a three base core sequence, a four base core sequence, or a five base core sequence. 2 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    [00010] In some embodiments, at least one of the bases of the core sequence comprises an arginine; a glycine and an arginine; a proline and an arginine; a lysine and an arginine; or any combination thereof. In some embodiments, the AT-rich interaction domain comprises a GRKPG (Gly-Arg-Lys-Pro-Gly) repeat, a RKRGRPKK repeat, a KKRGRPKK repeat, a RKRGR repeat, a GR*R/PPK repeat, a GR*RPK repeat, a GR*PPK repeat, a KRPR* repeat, or a K/RKRGRPKK repeat. In some embodiments, the AT-rich interaction domain comprises a core sequence comprising an amino acid selected from the group consisting of SEQ ID NO: 11- 24. [00011] In some embodiments of the engineered RT polypeptide described herein, the DNA binding domain is a DNA binding domain of any one of Saccharomyces cerevisiae datin (DAT1), high mobility group AT hook 1 (HMGA1), lysine-specific methyltransferase 2a ( KMT2A), Myocyte Enhancer Factor 2C (MEF2C), Heterogeneous Nuclear Ribonucleoprotein D (HNRNPD), Structural Maintenance of Chromosomes 1A (SMC1), Structural Maintenance Of Chromosomes 2 (SMC2), Caenorhabditis elegans tbp-1, Drosophila melanogaster D1 protein, Salmonella typhimurium Hin recombinase, S. typhimurium Gin recombinase, S. typhimurium Pin recombinase, or S. typhimurium Cin recombinase, or a combination thereof. [00012] In some embodiments, the DNA binding domain is from a S. cerevisiae DAT1. In some embodiments, the amino acid sequence of the DNA binding domain comprises a DNA binding domain consensus motif set forth in SEQ ID NO: 13, 14, 16, or 22. [00013] In some embodiments, the DNA binding domain comprises: (a) a full-length DAT1 sequence or SEQ ID NO: 2; (b) an N-terminal truncated variant of DAT1 (D90) comprising the first 90 amino acids of the full length DAT1, or SEQ ID NO: 3; (c) a truncated variant of DAT1 (D60) comprising the first 60 amino acids of the full length DAT1 or SEQ ID NO: 5; (d) a truncated variant of DAT1 (D48) comprising the first 48 amino acids of the full length DAT1 or SEQ ID NO: 6; (e) a truncated variant of DAT1(D36) comprising the first 36 amino acids of the full length DAT1 or SEQ ID NO: 8; (f) a truncated variant of DAT1 (D35) comprising the first 35 amino acids of full length DAT1 or SEQ ID NO: 9; (g) an amino acid sequence having at least about 90%, at least about 91%, at least about 92%, at least about 93%, at least about 94%, at least about 95%, at least about 96%, at least about 97%, at least about 98%, or at least about 99% sequence identity to any one of SEQ ID NO: 2, 3, 5, 6, 8, or 9; or (h) an amino acid 3 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    sequence having 90%, 91%, 92%, 93%, 94%, 95%, 96%, 97%, 98%, or 99% sequence identity to any one of SEQ ID NO: 2, 3, 5, 6, 8, or 9. [00014] In some embodiments, the DNA binding domain comprises the amino acid sequence of SEQ ID NO: 11 (GRKPG). In some embodiments, the DNA binding domain optionally comprise at least 2 domains or at least 3 domains comprising SEQ ID NO: 11. [00015] In some embodiments, the DNA binding domain comprises a mutation in any of one of SEQ ID NO: 2, 3, 5, 6, 8, 9, or 11. In some embodiments, the mutation is selected from a substitution, an insertion, a deletion, or any combination thereof. [00016] In some embodiments of the engineered RT polypeptide described herein, the DNA binding domain comprises SEQ ID NO: 2. In some embodiments of the engineered RT polypeptide described herein, the DNA binding domain comprises the amino acid sequence of SEQ ID NO: 3, 8, or 9. [00017] In some embodiments of the engineered RT polypeptide described herein, the RT polypeptide sequence comprises the amino acid sequence of SEQ ID NO: 7, and further comprises a combination of mutations selected from the group consisting of: (i) E69K, L139P, E302R, T306K, W313F, T330P, and N454K; and additionally one or more of M39V, P47L, Q91R, M66L, F155Y, D200N, D200E, H204R, G429S, L435G, L435K, P448A, D449G, H503V, D524N, T542D, E545G, D583N, H594Q, L603W, L603F, E607K, E607G, P627S, H634Y, H638G, A644V, D653H, K658R and L671P; and (ii) E69K, L139P, D200N, E302R, T306K, W313F, T330P, L435G, P448A, D449G, N454K, D524N, L603W, and E607K; and additionally one or more of M39V, P47L, M66L, Q91R, F155Y, H204R, G429S, H503V, T542D, E545G, D583N, H594Q, P627S, H634Y, H638G, A644V, D653H, K658R and L671P. [00018] In some embodiments of the engineered RT polypeptide described herein, the amino acid sequence of the RT polypeptide sequence is: (a) at least 90% identical to SEQ ID NO: 1 or 143; (b) about 90% to about 99.99% identical to SEQ ID NO: 1 or 143, about 92% to about 99.99% identical to SEQ ID NO: 1 or 143, about 93% to about 99.99% identical to SEQ ID NO: 1 or 143, about 94% to about 99.99% identical to SEQ ID NO: 1 or 143, about 95% to about 99.99% identical to SEQ ID NO: 1 or 143, about 96% to about 99.99% identical to SEQ ID NO: 1 or 143, about 97% to about 99.99% identical to SEQ ID NO: 1 or 143, or about 98% to about 4 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    99.99% identical to SEQ ID NO: 1 or 143; or (c) about 90%, about 91%, about 92%, about 93%, about 94%, about 95%, about 96%, about 97%, about 98%, about 99% or about 99.5% identical to SEQ ID NO: 1 or 143. [00019] In some embodiments, the RT polypeptide sequence comprises: (a) an amino acid sequence that is at least 95% identical to SEQ ID NO:1, 7, or 179, and; (b) a combination of mutations indexed to SEQ ID NO:7 or 178 selected from the group consisting of: (i) a combination of variants consisting of a T542D mutation, a D583N mutation, an E607G mutation, an A644V mutation, a D653H mutation, and a K658R mutation; and (ii) a combination of variants consisting of an E545G mutation, a D583N mutation, an H594Q mutation, an L603F mutation, and a S679P mutation. [00020] In some embodiments, the amino acid sequence of the RT polypeptide sequence comprises E69K, L139P, D200N, E302R, T306K, W313F, T330P, N454K, H503V, D524N, L603W, E607K, and H634Y. In some embodiments, the amino acid variations are at any one position or combination thereof as identified in an alignment of SEQ ID NO: 1 or 143 to any one of the RT polypeptide sequences in Table 1 or Table 2. [00021] In some embodiments, the RT polypeptide sequence comprises: (a) M39V, M66I, Q91R, I347V, and H594Q substitution in SEQ ID NO: 143; or (b) SEQ ID NO: 129 (SOLD 034). In some embodiments, the RT polypeptide sequence comprises M39V, T542D, D583N, E607G, A644V, D653H, K658R, and L671P in SEQ ID NO: 143. In some embodiments, the RT polypeptide sequence comprises: (a) M39V, T542D, D583N, E607G, A644V, D653H, K658R, L671P in SEQ ID NO: 143; or (b) SEQ ID NO: 111 (SOLD 025). In some embodiments, the RT polypeptide sequence comprises T542D, D583N, E607G, A644V, D653H, K658R, E545G, D583N, H594Q, and a L603F in SEQ ID NO: 143. [00022] One aspect of the present disclosure provides an engineered RT polypeptide comprising: (a) an amino acid sequence that is at least 90%, at least 92%, at least 95%, at least 97%, at least 98%, or at least 99% identical to: (i) an amino acid sequence of an RT disclosed in Table 1, or Table 2; or (ii) SEQ ID NOs: 27-61, 63, 65, 67, 69, 71, 73, 75, 77, 79, 81, 83, 85, 87, 89, 91, 93, 95, 97, 99, 101, 103, 105, 107, 109, 111, 113, 115, 117, 119, 121, 123, 125, 127, 129, 131, 133, 141, 143, 145, 147, 149, 151, 157, 159, 172, or 173; and (b) a DNA binding domain comprising an amino acid selected from the group consisting of SEQ ID NO: 2, 3, 5, 6, 8, and 9. 5 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    [00023] Another aspect of the present disclosure provides an engineered RT polypeptide comprising: (a) an amino acid sequence of an RT disclosed in Table 1 or Table 2; and (b) an amino acid sequence of DNA binding domain disclosed in Table 1. [00024] In some embodiments, the engineered RT polypeptide comprises: (a) the amino acid sequence of any one of SEQ ID NO: 174-188; (b) an amino acid sequence having at least 90%, at least 91%, at least 92%, at least 93%, at least 94%, at least 95%, at least 96%, at least 97%, at least 98%, or at least 99% sequence identity to any one of SEQ ID NO: SEQ ID NO: 174-188; or (c) an amino acid sequence having 90%, 91%, 92%, 93%, 94%, 95%, 96%, 97%, 98%, or 99% sequence identity to any one of SEQ ID NO: SEQ ID NO: 174-188. [00025] In some embodiments, the engineered RT comprises an amino acid sequence that is at least about 90% identical to an amino acid sequence selected from the group consisting of SEQ ID NO: 27-61, 63, 65, 67, 69, 71, 73, 75, 77, 79, 81, 83, 85, 87, 89, 91, 93, 95, 97, 99, 101, 103, 105, 107, 109, 111, 113, 115, 117, 119, 121, 123, 125, 127, 129, 131, 133, 141, 143, 145, 147, 149, 151, 157, 159, 172, and 173. [00026] In some embodiments, the RT polypeptide is 42B L (SEQ ID NO: 145), 50A+G (SEQ ID NO: 147), SOLD 022 (SEQ ID NO: 105), SOLD 023 (SEQ ID NO: 107), SOLD 025 (SEQ ID NO: 111), SOLD 031 (SEQ ID NO: 123), SOLD 033 (SEQ ID NO: 127), SOLD 034 (SEQ ID NO: 129), SOLD 035 (SEQ ID NO: 131), SOLD 001 (SEQ ID NO: 65), and SOLD 33 VDG (SEQ ID NO: 173), or an RT polypeptide set forth in SEQ ID NO: 143, or SEQ ID NO: 172. [00027] In some embodiments of the engineered RT polypeptide described herein, the engineered RT comprises at least two DNA binding domains. In some embodiments, at least one DNA binding domain is located at the N-terminus of the engineered RT and at least one DNA binding domain is located at the C-terminus of the engineered RT. In some embodiments, the at least two DNA binding domains are both located at the C-terminus or N-terminus of the engineered RT. [00028] One aspect of the present disclosure provides a recombinant reverse transcriptase (RT) protein comprising a RT polypeptide, fused to a DNA binding domain, where (a) the RT polypeptide and the DNA binding domain are separated by an amino acid linker, and (b) the DNA binding domain is fused to the C-terminus of the RT polypeptide. 6 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    [00029] One aspect of the present disclosure provides a recombinant reverse transcriptase (RT) protein comprising a RT polypeptide, fused to a DNA binding domain, where (a) the RT polypeptide and the DNA binding domain are separated by an amino acid linker, and (b) the DNA binding domain is fused to the N-terminus of the RT polypeptide. [00030] In some embodiments, the RT polypeptide is any one of the RT polypeptides listed in Table 1 or Table 2. [00031] In some embodiments of the recombinant reverse transcriptase (RT) protein described herein, the DNA binding domain is a DNA binding protein selected from the group consisting of S. cerevisiae datin (DAT1); high mobility group AT hook 1 (HMGA1), lysine-specific methyltransferase 2a ( KMT2A), Myocyte Enhancer Factor 2C (MEF2C), Heterogeneous Nuclear Ribonucleoprotein D (HNRNPD), Structural Maintenance of Chromosomes 1A (SMC1), Structural Maintenance Of Chromosomes 2 (SMC2), C. elegans tbp-1, D. melanogaster D1 protein, Salmonella typhimurium Hin recombinase, S. typhimurium Gin recombinase, S. typhimurium Pin recombinase, or S. typhimurium Cin recombinase. [00032] In some embodiments, the linker is a G(n)S(m)G(p) linker, where: (a) n = 0, 1, 2, 3, 4, 5,6 , 7, 8, 9, 1011, 12, 13, 14, 15, 16, 17, 18, 19 or 20; (b) m = 0, 1, 2, 3, 4, 5,6 , 7, 8, 9, 1011, 12, 13, 14, 15, 16, 17, 18, 19 or 20; (c) p = 0, 1, 2, 3, 4, 5, 6, 7, 8, 9, 1011, 12, 13, 14, 15, 16, 17, 18, 19 or 20; and (d) n, m, and p are selected independently. In some embodiments, the linker is a glycine-serine (GS) linker selected from the group consisting of (GS)n, (GSGGS)n, (SGGSG)n, (GGGS)n, (GGSG)n, (GGSGG)n, (GSGSG)n, (GSGGG)n, GGGSG)n, and (GSSSG)n, where n represents an integer of at least 1. In some embodiments, the linker is GGGS. In some embodiments, the linker is SGGSG. [00033] In some embodiments, the DNA binding domain is a S. cerevisiae datin (DAT1) DNA binding domain or fragment thereof. In some embodiments, the DNA binding domain comprises: (a) an amino acid sequence selected from the group consisting of SEQ ID NO: 2, 3, 5, 6, 8, 9, and 11-24; or (b) a nucleic acid sequence of SEQ ID NO: 25. [00034] In some embodiments, the RT polypeptide is selected from the group consisting of 42B L (SEQ ID NO: 145), 50A+G (SEQ ID NO: 147), SOLD 022 (SEQ ID NO: 105), SOLD 023 (SEQ ID NO: 107), SOLD 025 (SEQ ID NO: 111), SOLD 031 (SEQ ID NO: 123), SOLD 033 7 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    (SEQ ID NO: 127), SOLD 034 (SEQ ID NO: 129), SOLD 035 (SEQ ID NO: 131), SOLD 001 (SEQ ID NO: 65), SOLD 33 VDG (SEQ ID NO: 173), and an RT polypeptide set forth in SEQ ID NO: 143, SEQ ID NO: 172. [00035] Another aspect of the present disclosure provides a recombinant RT protein as described herein comprising, consisting essentially of, or consisting of SEQ ID NO: 174-188. [00036] In some embodiments of the engineered RT polypeptide described herein, or the recombinant RT protein described herein, the engineered RT polypeptide or the recombinant RT protein further comprises a tag protein selected from the group consisting of an affinity tag, a fluorescent tag, or an expression and/or solubility enhancement tag. [00037] In some embodiments of the engineered RT polypeptide or the recombinant RT protein described herein, the tag is selected from hexahistidine tag (his-tag), small ubiquitin-like modifier tag (SUMO), a VariFlex C-Terminal solubility enhancement tag, a short peptide C- terminal tag, Thioredoxin (Trx) tag, Solubility-enhancer peptide sequences (SET) tag, IgG domain B1 of Protein G (GB1) tag, IgG repeat domain ZZ of Protein A (ZZ) tag, Solubility enhancing Ubiquitous Tag (SNUT tag), Seventeen kilodalton protein (Skp tag), Phage T7 protein kinase (T7PK) tag, E. coli secreted protein A (EspA) tag, Monomeric bacteriophage T70.3 protein (Orc protein) (Mocr) tag, E. coli trypsin inhibitor (Ecotin) tag, Calcium-binding protein (CaBP) tag, Stress-responsive arsenate reductase (ArsC) tag, N-terminal fragment of translation initiation factor IF2 (IF2-domain I) tag, N-terminal fragment of translation initiation factor IF2 (Expressivity) tag, Fasciola hepatica 8-kDa antigen tag (Fh8), Glutathione-S-transferase (GST) tag, maltose-binding protein tag (MBP), Flag tag peptide (FLAG), streptavidin binding peptide tag (Strep-II; strep), calmodulin-binding protein tag (CBP), mutated dehalogenase tag (HaloTag), staphylococcal Protein A (Protein A), intein mediated purification with the chitin-binding domain (IMPACT (CBD)), cellulose-binding module (CBM), dockerin domain of Clostridium josui tag (Dock), fungal avidin-like protein (Tamavidin). [00038] In some embodiments of the engineered RT polypeptide or the recombinant RT protein described herein, the tag is an affinity tag selected from hexahistidine tag (his-tag), Fasciola hepatica 8-kDa antigen tag (Fh8), Glutathione-S-transferase (GST) tag, maltose-binding protein tag (MBP), Flag tag peptide (FLAG), streptavidin binding peptide tag (Strep-II), calmodulin- binding protein tag (CBP), mutated dehalogenase tag (HaloTag), staphylococcal Protein A 8 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    (Protein A), intein mediated purification with the chitin-binding domain (IMPACT (CBD)), cellulose-binding module (CBM), dockerin domain of Clostridium josui tag (Dock), fungal avidin-like protein (Tamavidin). [00039] In some embodiments of the engineered RT polypeptide or the recombinant RT protein described herein, the engineered RT polypeptide or the recombinant RT protein comprises: (a) an hexahistidine tag (his-tag); or (b) an amino acid sequence of SEQ ID NO: 62; or an amino acid sequence having at least 90% sequence identity to the amino acid sequence of SEQ ID NO: 62. [00040] In some embodiments of the engineered RT polypeptide or the recombinant RT protein described herein, the engineered RT polypeptide or the recombinant RT protein comprises a solubility enhancer tag selected from the group consisting of a SUMO tag, a GST tag, a Trx tag, a VariFlex C-Terminal solubility enhancement tag, a short peptide C-terminal tag, an Fh8 tag, MBP tag, SET tag, GB1 tag, ZZ tag, HaloTag, SNUT tag, Skp tag, T7PK tag, EspA tag, Mocr tag, Ecotin tag, CaBO tag, ArsC tag, IF2-domain I tag, Expressivity tag, RpoA, tag, SlyD, tag, Tsf tag, RpoS tag, PotD tag, Crr tag, msyB tag, yigD tag, and rpoD tag. [00041] In some embodiments of the engineered RT polypeptide or the recombinant RT protein described herein, the engineered RT polypeptide or the recombinant RT protein comprises: (a) a short peptide C-terminal tag; (b) an amino acid sequence of SEQ ID NO: 193; or (c)an amino acid sequence having at least 90% sequence identity to the amino acid sequence of SEQ ID NO: 193. [00042] In some embodiments of the engineered RT polypeptide or the recombinant RT protein described herein, the tag further comprises: (a) an endoprotein cleavage sequence; (b) a cleavage sequence recognized by an endoprotein selected from the group consisting of alanine carboxypeptidase, Armillaria mellea astacin, bacterial leucyl aminopeptidase, cancer procoagulant, cathepsin B, clostripain, cytosol alanyl aminopeptidase, elastase, endoproteinase Arg-C, enterokinase (EnTK), gastricsin, gelatinase, Gly-X carboxypeptidase, glycyl endopeptidase, human rhinovirus 3C protease, hypodermin C, Iga-specific serine endopeptidase, leucyl aminopeptidase, leucyl endopeptidase, lysC, lysosomal pro-X carboxypeptidase, lysyl aminopeptidase, methionyl aminopeptidase, myxobacter, nardilysin, pancreatic endopeptidase E, picornain 2A, picornain 3C, proendopeptidase, prolyl aminopeptidase, proprotein convertase I, proprotein convertase II, russellysin, saccharopepsin, semenogelase, T-plasminogen activator, 9 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    thrombin (Thr), tissue kallikrein, tobacco etch virus (TEV), togavirin, tryptophanyl aminopeptidase, U-plasminogen activator, V8, venombin A, venombin AB, factor Xa (Xa), and Xaa-pro aminopeptidase; or (c) an endoprotein cleavage sequence comprising the amino acid sequence of SEQ ID NO: 194, SEQ ID NO: 195, SEQ ID NO: 196, SEQ ID NO: 197, or SEQ ID NO: 198. [00043] In some embodiments of the engineered RT polypeptide described herein, or the recombinant RT protein described herein, the engineered RT polypeptide or the recombinant RT protein exhibits increased template switching (TS) efficiency, increased processivity efficiency, increased binding affinity, increased transcription efficiency, increased chemical tolerance, improved ability to yield mitochondrial unique molecular identity (UMI) counts, improved ability to yield ribosomal unique molecular identity (UMI) counts, longer shelf life, higher strand displacement, higher end-to-end template jumping, or any combination thereof, when compared to an RT polypeptide or a recombinant RT protein lacking a conjugated DNA binding domain. [00044] In some embodiments of the engineered RT polypeptide described herein, or the recombinant RT protein described herein, the engineered RT polypeptide or the recombinant RT protein comprises at least two or more of increased template switching (TS) efficiency, increased processivity efficiency, increased binding affinity, increased transcription efficiency, increased chemical tolerance, improved ability to yield mitochondrial unique molecular identity (UMI) counts, longer shelf life, higher strand displacement, higher end-to-end template jumping, or improved ability to yield ribosomal unique molecular identity (UMI) counts, when compared to an RT polypeptide or a recombinant RT protein lacking a conjugated DNA binding domain. [00045] In some embodiments of the engineered RT polypeptide described herein or the recombinant RT protein described herein, the recombinant RT protein or the engineered RT exhibits increased transcript capture during amplification, when compared to an RT polypeptide or a recombinant RT protein lacking a conjugated DNA binding domain. [00046] In some embodiments of the engineered RT polypeptide or the recombinant RT protein described herein, the DNA binding domain enhances the hybridization of a transcript and a primer during a nucleic acid amplification process, when compared to an RT polypeptide or a recombinant RT protein lacking a conjugated DNA binding domain. 10 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    [00047] In some embodiments of the engineered RT polypeptide or the recombinant RT protein described herein, the primer comprises a poly-dT or a poly(dT)VN sequence and a non-poly(dT) sequence; and the transcript comprises a poly-dA sequence. [00048] In some embodiments of the engineered RT polypeptide or the recombinant RT protein described herein, the DNA binding domain stabilizes the oligo(A)-oligo(T) based transcript- primer complex during a nucleic acid amplification process. [00049] In some embodiments of the engineered RT polypeptide or the recombinant RT protein described herein, the primer is a barcoded molecule. [00050] In some embodiments of the engineered RT polypeptide or the recombinant RT protein described herein, the transcript is a nucleic acid molecule selected from an RNA, a mRNA, or a DNA. [00051] One aspect of the present disclosure provides an isolated nucleic acid molecule encoding: (a) an engineered RT polypeptide described herein; or (b) a recombinant RT protein described herein. [00052] In some embodiments, the nucleic acid molecule comprises a sequence selected from SEQ ID NO: 25, SEQ ID NO: 136, SEQ ID NO:138, SEQ ID NO:140, SEQ ID NO:142, SEQ ID NO:144, SEQ ID NO:146, SEQ ID NO:148, SEQ ID NO:150, SEQ ID NO:152, SEQ ID NO:154, SEQ ID NO:156, SEQ ID NO:158, SEQ ID NO:160, SEQ ID NO:162, SEQ ID NO:164, SEQ ID NO:167, SEQ ID NO:169, or SEQ ID NO: 171; or a nucleic acid sequence of Table 2. [00053] One aspect of the present disclosure provides an expression vector comprising any isolated nucleic acid described herein. [00054] One aspect of the present disclosure provides a host cell transfected with any expression vector described herein or any isolated nucleic acid described herein. [00055] One aspect of the present disclosure provides a composition comprising: (a) any recombinant RT protein described herein; or (b) any engineered RT polypeptide described herein; or (d) any expression vector described herein; or (e) any host cell described herein; and (f) a buffer. 11 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    [00056] One aspect of the present disclosure provides a method for performing a reverse transcription reaction for generating a nucleic acid product from an RNA template comprising contacting under suitable conditions a biological sample or extract thereof with an engineered RT polypeptide described herein, or a recombinant RT protein describe herein. [00057] In some embodiments, the biological sample or extract thereof comprises a cell, optionally the cell is permeabilized and/or optionally the cell is fixed. In some embodiments, the biological sample or extract thereof comprises a cell bead, optionally the cell bead is fixed. In some embodiments, the biological sample or extract thereof comprises a nucleus, optionally the nucleus is permeabilized and optionally the nucleus is fixed. In some embodiments, the biological sample or extract thereof comprises (a) a suitable cellular preparation selected from cell populations and/or single cells, or (b) a tissue. [00058] In some embodiments, the biological sample or extract thereof comprises: (a) a cell, a cell bead, a permeabilized cell, a nucleus, where the nucleus is optionally permeabilized, and/or optionally the cell, the cell bead, the permeabilized cell and/or nucleus are fixed; (b) a suitable cellular preparation selected from cell populations and/or single cells; or (c) a tissue. [00059] In some embodiments, the biological sample or extract thereof comprises cells in suspension, fresh cells, fixed cells, or cells and tissues immobilized on various solid surfaces. [00060] In some embodiments, when the biological sample is a cell, a cell bead, or a nucleus, the reverse transcription reaction is part of a single cell RNA sequencing assay. In one embodiment, the single cell RNA sequencing assay further comprises, prior to the reverse transcription, partitioning the cell, the cell bead, or the nucleus into a partition. In one embodiment, the single cell RNA sequencing assay further comprises, after the reverse transcription reaction, hybridizing the nucleic acid product to an oligonucleotide molecule comprising a partition-specific barcode. [00061] In some embodiments, when the biological sample is a cell or tissue sample immobilized on a surface, the reverse transcription reaction is part of a spatial RNA sequencing assay. [00062] In some embodiments of the method described herein, the engineered RT polypeptide or the recombinant RT protein enhances template switching (TS) efficiency, processivity 12 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    efficiency, binding affinity, transcription efficiency, chemical tolerance, ability to yield mitochondrial unique molecular identity (UMI) counts, ability to yield ribosomal unique molecular identity (UMI) counts, strand displacement, end-to-end template jumping, or any combination thereof, when compared to an RT polypeptide or a recombinant RT protein lacking a conjugated DNA binding domain. [00063] In some embodiments, the engineered RT polypeptide or the recombinant RT protein enhances at least two or more of template switching (TS) efficiency, processivity efficiency, binding affinity, transcription efficiency, chemical tolerance, ability to yield mitochondrial unique molecular identity (UMI) counts, strand displacement, end-to-end template jumping, or ability to yield ribosomal unique molecular identity (UMI) counts, when compared to an RT polypeptide or a recombinant RT protein lacking a conjugated DNA binding domain. [00064] In some embodiments, the recombinant RT protein or the engineered RT enhances transcript capture during amplification, when compared to an RT polypeptide or a recombinant RT protein lacking a conjugated DNA binding domain. [00065] In some embodiments, the DNA binding domain of the recombinant RT protein or the engineered RT enhances the hybridization of a transcript and a primer during the amplification process, when compared to an RT polypeptide or a recombinant RT protein lacking a conjugated DNA binding domain. [00066] In some embodiments, the engineered RT polypeptide or the recombinant RT protein comprises: (a) a DNA binding domain comprising an amino acid sequence selected from SEQ ID NO: 2, 3, 5, 6, 8, 9, or 11-24; and (b) an amino acid sequence selected from SEQ ID NOs: 27- 61, 63, 65, 67, 69, 71, 73, 75, 77, 79, 81, 83, 85, 87, 89, 91, 93, 95, 97, 99, 101, 103, 105, 107, 109, 111, 113, 115, 117, 119, 121, 123, 125, 127, 129, 131, 133, 141, 143, 145, 147, 149, 151, 157, 159, 172, or 173. [00067] In some embodiments of the method described herein, the amino acid sequence of the engineered RT polypeptide or the recombinant RT protein comprises an amino acid sequence having at least about 90%, at least about 92%, at least about 95%, at least about 97%, at least about 98%, or at least about 99% sequence identity to SEQ ID NO: 174-188. 13 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    [00068] In some embodiments of the method described herein, the engineered RT polypeptide or the recombinant RT protein comprises: (a) M39V, M66I, Q91R, I347V, and H594Q substitution in SEQ ID NO: 143; (b) SEQ ID NO: 129 (SOLD 034); (c) SOLD 001 (SEQ ID NO: 65); or (d) SOLD 33 VDG (SEQ ID NO: 173). [00069] In some embodiments, the engineered RT polypeptide or the recombinant RT protein comprises M39V, T542D, D583N, E607G, A644V, D653H, K658R, and L671P in SEQ ID NO: 1 or 143. In some embodiments, the engineered RT polypeptide or the recombinant RT protein comprises: (a) M39V, T542D, D583N, E607G, A644V, D653H, K658R, L671P in SEQ ID NO: 143; or (b) SEQ ID NO: 111 (SOLD 025). [00070] In some embodiments, the engineered RT or the recombinant RT protein comprises a M39V, M66I, Q91R, I347V, H594Q in SEQ ID NO: 143. In some embodiments, the engineered RT polypeptide or the recombinant RT protein comprises an amino acid sequence that is at least about 90%, at least about 92%, at least about 95%, at least about 97%, at least about 98%, or at least about 99% identical to an amino acid sequence disclosed in Table 1 or Table 2. [00071] One aspect of the present disclosure provides a method of using an engineered RT polypeptide described herein, or a recombinant RT protein described herein, the method comprising contacting the engineered RT polypeptide or the recombinant RT protein with a nucleic acid template under suitable conditions to produce a polymerized nucleic acid product, where the nucleic acid template comprises an RNA, a DNA, or a nucleic acid comprising an unnatural nucleotide. In some embodiments, the nucleic acid template comprises an RNA. [00072] One aspect of the present disclosure provides a nucleic acid extension method comprising: (a) contacting a target nucleic acid molecule with an engineered reverse transcriptase polypeptide or a recombinant RT protein and a plurality of nucleic acid barcoded molecules comprising a barcode sequence, and (b) incubating the target nucleic acid, the engineered RT polypeptide or the recombinant RT protein and barcoded molecules under suitable conditions in which the barcoded molecules are extended by the engineered RT polypeptide or the recombinant RT protein. In some embodiments, the engineered RT polypeptide comprises the amino acid sequence of an engineered RT polypeptide described herein, or a recombinant RT protein described herein. 14 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    [00073] In some embodiments, the recombinant RT protein or the engineered RT polypeptide exhibits increased transcript capture during amplification. In some embodiments, the DNA binding domain enhances the hybridization of a transcript and a primer during a nucleic acid amplification process. In some embodiments, the primer comprises a poly-dT or a poly(dT)VN sequence and a non-poly(dT) sequence; and the transcript comprises a poly-dA sequence. In some embodiments, the DNA binding domain stabilizes the oligo(A)-oligo(T) based transcript- primer complex during a nucleic acid amplification process. In some embodiments, the primer is a barcoded molecule. [00074] In some embodiments, the recombinant RT protein or the engineered RT polypeptide described herein performs the first strand complementary DNA (cDNA) reaction. In that embodiment, the first strand cDNA is amplified using a DNA polymerase to generate a second strand cDNA. [00075] One aspect of the present disclosure provides a method of producing an engineered RT polypeptide or recombinant RT protein of the present disclosure, the method comprising providing a composition comprising a cell lysate and/or cellular fraction comprising the engineered RT polypeptide or the recombinant RT protein and subjecting the composition to protein purification steps so as to produce a substantially purified engineered RT enzyme. [00076] One aspect of the present disclosure provides a kit comprising: (a) a recombinant RT protein described herein; or (b) an engineered reverse transcriptase polypeptide described herein; or (c) the isolated nucleic acid described herein; or (d) an expression vector described herein; or (e) a host cell described herein; or (f) a composition described herein; and (g) instructions. [00077] Both the foregoing summary and the following description of the drawings and detailed description are exemplary and explanatory. They are intended to provide further details of the invention but are not to be construed as limiting. Other objects, advantages, and novel features will be readily apparent to those skilled in the art from the following detailed description of the invention. [00078] Section headings, numerical and/or alphabetical listings, e.g., (a), (b), (i) etc., are presented merely for ease of reading the disclosure, including the specification and claims. The use of headings in the disclosure, including the specification or claims does not require the steps 15 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    or elements be performed in alphabetical or numerical order or the order in which they are presented. [00079] Various embodiments of the features of this disclosure are described herein. However, it should be understood that such embodiments are provided merely by way of example, and numerous variations, changes, and substitutions can occur to those skilled in the art without departing from the scope of this disclosure. It should also be understood that various alternatives to the specific embodiments described herein are also within the scope of this disclosure. BRIEF DESCRIPTION OF THE DRAWINGS [00080] The following drawings illustrate certain embodiments of the features and advantages of this disclosure. These embodiments are not intended to limit the scope of the appended claims in any manner. Like reference symbols in the drawings indicate like elements. [00081] FIG.1A shows an exemplary sandwiching process where a first substrate (e.g., a slide), including a biological sample, and a second substrate (e.g., array slide) are brought into proximity with one another. [00082] FIG.1B shows a fully formed sandwich configuration creating a chamber formed from the one or more spacers, the first substrate, and the second substrate. [00083] FIG.2A shows a perspective view of an exemplary sample handling apparatus in a closed position. [00084] FIG.2B shows a perspective view of an exemplary sample handling apparatus in an open position. [00085] FIG.3A shows the first substrate angled over (superior to) the second substrate. [00086] FIG.3B shows that as the first substrate lowers, and/or as the second substrate rises, the dropped side of the first substrate may contact a drop of reagent medium. [00087] FIG.3C shows a full closure of the sandwich between the first substrate and the second substrate with one or more spacers contacting both the first substrate and the second substrate. [00088] FIG.4A shows a side view of the angled closure workflow. [00089] FIG.4B shows a top view of the angled closure workflow. 16 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    [00090] FIG.5 is a schematic diagram showing an example of a barcoded capture probe, as described herein. [00091] FIG.6 shows a schematic illustrating a cleavable capture probe. [00092] FIG.7 shows exemplary capture domains on capture probes. [00093] FIG.8 shows an exemplary arrangement of barcoded features within an array. [00094] FIG.9A shows and exemplary workflow for performing a templated capture and producing a ligation product, and FIG.9B shows an exemplary workflow for capturing a ligation product from FIG.9A on a substrate. [00095] FIG.10 is a schematic diagram of an exemplary analyte capture agent. [00096] FIG.11 is a schematic diagram depicting an exemplary interaction between a feature- immobilized capture probe 1124 and an analyte capture agent 1126. [00097] FIG.12 shows a schematic diagram of a non-limiting embodiment of a generalized capture probe used in spatial transcriptomics and/or single cell transcriptomic analyses, exemplary applications in addition to general reverse transcription reactions where the engineered reverse transcriptase of the disclosure could be used to extend a capture probe using a captured target nucleic acid as a template, thereby generating a cDNA product. [00098] FIG.13 provides a schematic of an exemplary capillary electrophoresis (CE) validation assay process used to test the activity of candidate enzymes. In step 1, 5’-end labeled DNA primers were hybridized to RNA templates at room temperature (approx.25°C); and poly rG- labeled template switching oligonucleotides (rG-TSO) were added to the reaction mixture. In step 2, the temperature was raised to about 53°C and first strand cDNA was synthesized with the addition of a poly-C tail (tailing). In step 3, template switching and TSO extension were performed. In step 4, the amplification product was transferred to a Genetic Analyzer for analysis. [00099] FIGs.14A-B provide schematics of an exemplary single cell and spatial assay for transcript capture. FIG.14A illustrates a schematic process of the 5’ single cell assay and FIG. 14B illustrates a schematic process of the 3’ single cell assay and step 1 of Visium. The first step of both assays is the hybridization of an oligo(A) from an mRNA to an oligo(T)) of a primer and 17 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    binding of reverse transcriptase to the annealed primer-template (which contains an oligo(A)- oligo(T) tract). [000100] FIGs.15A-B provide schematics of exemplary Visium/ spatial 3’/5’ workflows. FIG. 15A illustrates a schematic of a Visium 3’ workflow demonstrating a polyA capture (box). FIG. 15B illustrates a schematic of a Visium ATP 5’ workflow demonstrating a polyA capture (box). [000101] FIGs.16A-B provide a non-limiting embodiment of a nucleic acid sequence (FIG. 16A) and an amino acid sequence (FIG.16B) of an exemplary engineered reverse transcriptase (RT) polypeptide described in the present disclosure (N-Dat_42BL) comprising a DNA binding domain (e.g., DAT; bold and underline) operably linked to a reverse transcriptase (42BL) via a linker (bold). [000102] FIGs.17A-B provide a non-limiting embodiment of a nucleic acid sequence (FIG. 17A) and an amino acid sequence (FIG.17B) of an exemplary engineered reverse transcriptase (RT) polypeptide described in the present disclosure (C-Dat_42BL) comprising a reverse transcriptase (42BL) operably linked to a DNA binding domain (e.g., DAT; bold and underline) via a linker (bold). [000103] FIGs.18A-B provide a non-limiting embodiment of a nucleic acid sequence (FIG. 18A) and an amino acid sequence (FIG.18B) of an exemplary engineered reverse transcriptase (RT) polypeptide described in the present disclosure (N-DAT1full_42BL) comprising a DNA binding domain (e.g., full length DAT1 protein; bold and underline) operably linked to a reverse transcriptase (42BL) via a linker (bold). [000104] FIGs.19A-B provide a non-limiting embodiment of a nucleic acid sequence (FIG. 19A) and an amino acid sequence (FIG.19B) of an exemplary engineered reverse transcriptase (RT) polypeptide described in the present disclosure (C-DAT1full_42BL) comprising a reverse transcriptase (42BL) operably linked to a DNA binding domain (e.g., full length DAT1 protein; bold and underline) via a linker (bold). [000105] FIG.20 shows the performance of two engineered RTs described herein (N-DAT 42BL; SEQ ID NO: 175) and C-DAT 42BL (SEQ ID NO: 174)) in a Single Cell 5’ (SC-5’) gene expression assay when compared to two control MMLV variants (SOLD 001 (SEQ ID NO: 65) and SOLD 33 VDG (SEQ ID NO: 175)); and illustrates the superiority of the DAT engineered 18 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    RTs single cell assays over control non-DAT engineered RTs based on the median genes and UMIs per cell at 20k and 50k raw-reads per cell (rrpc). [000106] FIG.21 shows the performance of two engineered RTs described herein (N-DAT 42BL; SEQ ID NO: 175) and C-DAT 42BL (SEQ ID NO: 174)) in a Single Cell 5’ (SC-5’) gene expression assay when compared to two control MMLV variants (SOLD 001 (SEQ ID NO: 65) and SOLD 33 VDG (SEQ ID NO: 175)); and illustrates the superiority of the DAT engineered RTs for single cell assays over control non-DAT engineered RTs based on spatial transcriptomics and single cell transcriptomic analyses. [000107] FIGs.22A-B provide a bar graph (FIG.22A) and a quantification (FIG.22B) of the relative differences in performance between three reverse transcriptases when compared to a control RT and illustrating that a clear performance gains can be seen for a DAT fusion on either the N-terminal or C-terminal domain. In particular, FIGs.22A-B show median genes and UMIs/cell at 50k raw-reads per cell of three reverse transcriptases with and without a DAT fusion domain. [000108] FIGs.23A-D provide graphs illustrating the performance of various engineered RTs disclosed herein at maximum normalization depth; and showing median genes (FIG.23A) and UMIs/cell (FIG.23B) at maximum normalized read depth comparing library complexity of three reverse transcriptases with and without the DAT fusion domain. FIGs.23C-D show saturation curves of the median genes (FIG.23C) and counts/cell (FIG.23C) as a function of read depth. At maximum normalized read depth, the benefit of the DAT fusion on each RT backbone can clearly be seen. [000109] FIGs.24A-F provide graphs illustrating the differential gene expression of some engineered RTs comprising DAT1 at the N-terminus. FIGs.24A, C, and E feature scatter plots showing gene expression correlation of three reverse transcriptases with and without the DAT fusion domain. FIGs.24B, D, and F feature volcano plots showing the number of differentially expressed genes between three reverse transcriptases with and without the DAT fusion domain. [000110] FIGs.25A-F provide graphs illustrating the differential gene expression of some engineered RTs comprising DAT1 at the C-terminus. FIGs.25A, C, and E feature scatter plots showing gene expression correlation comparison of engineered RTs comprising N-Terminal 19 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    DAT Domain and C-Terminal DAT Domain. FIGs.25B (42BL versus 42BL NDAT), D, and F feature volcano plots comparing the number of differentially expressed genes between various RTs. Gains in performance were still present but were not as significant as a C-terminal DAT fusion. For example, the C-terminal DAT fusion did not perform as well as the N-terminal DAT fusion, but both N-terminal DAT fusion and C-terminal DAT fusion showed performance gains when compared to a control RT variant. The control RT variant was a non-DAT fusion RT comprising the same RT backbone (e.g., RT backbone alone; SEQ ID NO: 145). [000111] FIGs.26A-D provide performance comparison of the impact of the DAT domain across three reverse transcriptase backbones. FIGs.26A-B summarize median genes (FIG.26A) and UMIs/cell (FIG.26B) at maximum normalization depth comparing the performance among three MMLV RT variants 42B, 42BL, and 50A+ G backbones with and without a N-terminal DAT fusion domain; and illustrate a clear performance benefit from the DAT domain. FIG.26C shows gene expression correlation. FIG.26D shows differential gene expression. MMLV RT variants without a DAT fusion domain were used as controls. [000112] FIG.27 provides a schematic of exemplary Visium/ spatial 3’ workflows highlighting the improvements associated with the engineered RT variant comprising DAT1 at the N- terminus (NDAT1) disclosed herein. Six areas of target optimization (stars) are shown when using the NDAT1 RT variant disclosed herein. During the permeabilization step, the NDAT1 RT variant reduced transcript mislocalization and increased target capture. Cleaving oligos after annealing and reducing steric hindrance also enhanced optimization. During reverse transcription, there was an improved template switching and increased RT efficiency. During the second strand synthesis, there was an improved synthesis efficiency and increased product. During the cDNA amplification, amplification was increased without additional noise, and enough sample for long reads was provided. During the repair, A-tailing, and ligation steps, repair and efficiency were improved and no data was lost. [000113] FIGs.28A-F provide UMI heat maps showing globally detected gene expression in two replicates of spatial assay performed using a first slide configuration, on a human tonsil tissue, using an NDAT1 MMLV RT variant (FIGs.28C-F) or a control MMLV RT variant (FIGs.28A-B) and three different buffer formulations – a commercially available RT reagent (FIG.28A and FIG.28C), Buffer X.4 (FIG.28D), and Buffer X.5 (FIG.28B, FIG.28E, and 20 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    FIG.28F). In particular, FIGs.28A-F show that NDAT1 RT variant improved the sensitivity of the assay while maintaining good spatial resolution when compared to the control RT variant, with all three buffer conditions. [000114] FIGs.29A-E provide graphs quantifying the quality, sensitivity, and detection of gene expression under the same six conditions shown in FIGs.28A-F. In particular, the following metrics were evaluated: fraction reads in spots under tissue (FIG.29A), reads mapped confidently to transcriptome (FIG.29B), total genes detected (FIG.29C), GRch38 median UMI counts per spot (30k mapped spot-reads per spot) (FIG.29D), and GRch38 median genes per spot (30k mapped spot-reads per spot) (FIG.29E). The quantification confirmed that NDAT1 RT variant improved the sensitivity of the assay while maintaining good spatial resolution when compared to the control RT variant (FIGs.29D-E). A MMLV RT variant without a DAT fusion domain was used as a control. [000115] FIGs.30A-F provide UMI heat maps showing gene expression of RGS3 in a human tonsil tissue analyzed using NDAT1 RT variant (FIGs.30C-F) or a control RT variant (FIGs. 30A-B) and three different buffer formulations - RT reagent (FIG.30A and FIG.30C), Buffer X.4 (FIG.30D), and Buffer X.5 (FIG.30B, FIG.30E, and FIG.30F). A MMLV RT variant without a DAT fusion domain was used as a control. Under sandwich configuration conditions, some evidence of undesirable flow of transcripts and/or target molecules or analytes (black circle on upper right quadrant indicates images with a tail) was observed. UMI heat maps are shown as a log10(UMI) ranging from cold to hot (0, 0.5, 1.0, 1.5, and 2.0). [000116] FIGs.31A-C provide UMI heat maps showing gene expression of KRT5 in two replicates of a spatial assay performed using a first slide configuration on human tonsil tissues analyzed using NDAT1 RT variant (FIGs.31C-F) or a control RT variant (FIGs.31A-B; SEQ ID NO: 1, 142, 143, or 172) and three different buffer formulations - RT reagent (FIG.31A and FIG.31C), Buffer X.4 (FIG.31D), and Buffer X.5 (FIG.31B, FIG.31E, and FIG.31F). A MMLV RT variant without a DAT fusion domain was used as a control. Under sandwich configuration conditions, some evidence of undesirable flow of transcripts and/or target molecules or analytes (black circle on upper right quadrant) was observed. UMI heat maps are shown as a log10(UMI) ranging from cold to hot (0, 0.5, 1.0, 1.5, 2.0, and 2.5). 21 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    [000117] FIGs.32A-E provide UMI heat maps showing globally detected gene expression (FIGs.32A-B) or gene expression of Crgm1 (FIG.32C), KRT5 (FIG.32D), and Rho (FIG. 32E) in a spatial assay performed using a first slide configuration on Zebrafish (non-human or mouse tissue) (e.g., Visium standard definition (SD) spatial gene expression) analyzed using NDAT1 RT variant (42BL-NDAT1) and RT reagent buffer under sandwich configuration conditions. UMI heat maps are shown as a count log10 ranging from cold to hot (0.5, 1, 1.5, 2, 2.5, 3, 4, 4.5, and 5), or log normalized per experiment for CRYGM1 (0.00-5.0), KRT5 (0.0-7.0) or RHO (0.0-8.0). The quality, and sensitivity metrics were: fraction reads in spots under tissue (0.9), fraction reads usable (0.4), fraction reads mapped to genome (0.56), and number of reads (250M). [000118] FIGs.33A-D provide UMI heat maps showing globally detected gene expression (FIG.33A) or gene expression of CRGM1 (FIG.33B), KRT5 (FIG.33C), and RHO (FIG. 33D) in a spatial assay performed using a second slide configuration (e.g., Visium high definition (HD) spatial gene expression or next generation spatial gene expression) on Zebrafish (non-human or mouse tissue) analyzed using NDAT1 RT variant (42BL-NDAT1) and RT reagent buffer under sandwich configuration conditions. UMI heat maps are shown as a log normalized per experiment for CRYGM1 (0.00-4.0), KRT5 (0.0-9.0) or RHO (0.0-7.0). The quality, and sensitivity metrics were fraction reads mapped to genome (0.72). DETAILED DESCRIPTION I. OVERVIEW A. Enhancement of Transcript Capture [000119] One goal of next generation single-cell and spatial platforms is to improve the sensitivity (or UMI and transcript capture) of single cell assays. An early, and pivotal step, of transcript capture in single cell 5', single cell 3', and spatial assays (i.e., Visium) is onboarding of reverse transcriptase enzyme to hybridized oligo(A)-oligo(T) primer-template tract, of which the poly-dT arises from the primer and the poly-A from the mRNA transcript. [000120] In current assays and in published protocols, methods to optimize transcript capture using sequence specific DNA binding proteins have not been described. In order to improve the transcript capture (and therefore sensitivity) of existing assays, the present disclosure leverages 22 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    functional characteristics of DNA binding proteins to engineer RT variants that increase the sensitivity of the RT in single cell applications. In particular, the present disclosure provides engineered RT variants (e.g., engineered reverse transcriptase (RT) polypeptide or recombinant RT proteins) to increase transcript capture during single application using an accessory protein that is sequence specific for oligo(A)-oligo(T) tract. [000121] The accessory protein contemplated by the present disclosure is a DNA binding protein or amino acid motif which is sequence specific to oligo(A)-oligo(T) tract. The accessory protein can function as a “fusion” partner with the reverse transcriptase (e.g., MMLV RT or variant thereof) in single cell 5', single cell 3', and/or spatial assays (e.g., 10X Genomics single cell 5', single cell 3', and/or spatial assays). This strategy enabled improvement of transcript capture, and therefore sensitivity of the assays. B. DAT1 [000122] Dating (DAT1) or a truncation thereof was identified as a candidate sequence specific DNA binding motif. DAT1 is a yeast protein (e.g., Saccharomyces cerevisiae) that specifically recognizes the minor groove of non-alternating oligo(A)-oligo(T) tracts (e.g., >10 bp oligo(A)- oligo(T) tract). See e.g., Reardon et al., PNAS 90, 11327 (1993); Reardon et al. Nucleic Acids Research, 23, 4900 (1995); and Winter & Varshavsky, EMBO J., 8:1867 (1989). The sequence specific recognition may be determined by three repeated pentads of G-R-K-P-G (SEQ ID NO: 11). The N-terminal 90 amino acids (D90) and/or the N-terminal 36 amino acids (D36) can bind in a sequence specific manner to oligo(A)-oligo(T) tract. DAT1(D-90) can specifically bind to A-T tracts with Kd of about 3 x 10-10 M (or 3 x 10-9 M); and DAT1(D-36) protein can bind to A- T tracts with Kd of 4 x 10-10 M. DAT1(D-90) can also be more resistant to degradation by bacterial proteases than longer and shorter DAT1 derivatives. The DNA binding activity of DAT1(D-90) can also be resistant to heat (boiling in water bath for 10 min) and chemical treatment (6 M guanidine HCl). DAT1(D-90) can also be highly soluble in physiologic salt and pH conditions. [000123] While DNA binding proteins have been used in combination with a reverse transcriptase as fusion proteins to improve processivity of the reverse transcriptase, sequence specific DNA binding proteins have not been explored. See e.g., Oscorbin et al., FEBS Lett., 594, 4338 (2020). Furthermore, sequence specific DNA binding proteins have not been applied 23 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    to sequencing applications of any type. Fusion proteins with DAT1 have also not been constructed. [000124] As such, engineered RT molecules comprising DAT1 or fragments thereof were investigated as methods to improve the sensitivity of single cell assays. The binding properties of DAT1 was found to be highly tunable. DAT1 was combined with reverse transcriptase (42B or other MMLV variants). Both N-terminal and C-terminal fusions were explored. In addition, different truncations of DAT1 were tested (i.e., DAT1(90), first 90 N-terminal amino acids; or DAT1(36), a 36 residue minimal DAT1 binding domain). The present disclosure shows that DAT1 binding domain can assist a reverse transcriptase in binding to primed transcripts (FIGs. 14A-B and 15A-B; box), thereby increasing the assay sensitivity. The DAT1 constructs were optimized for: (1) truncation of DAT1 binding domain; (2) tuning the binding strength of DAT1 to oligo(A)-oligo(T) tract through sequence modification (e.g., removal or alteration of G-R-K- P-G binding pentad); and (3) identification and testing of other sequence-specific DNA binding proteins. C. Experimental results [000125] Accordingly, the present disclosure provides engineered reverse transcriptase (RT) polypeptides comprising an RT polypeptide sequence; a DNA binding domain; and a linker connecting the RT polypeptide sequence and the DNA binding domain. The DNA binding domain is from a molecule capable of binding a minor groove of a nucleic acid (e.g., DAT 1). In some aspect, the present disclosure provides recombinant reverse transcriptase (RT) proteins comprising a RT polypeptide, fused to a DNA binding domain, where the RT polypeptide and the DNA binding domain are separated by an amino acid linker, and the DNA binding domain is fused to the N-terminus of the RT polypeptide. In another aspect, the present disclosure provides recombinant reverse transcriptase (RT) proteins comprising a RT polypeptide, fused to a DNA binding domain, where the RT polypeptide and the DNA binding domain are separated by an amino acid linker, and the DNA binding domain is fused to the C-terminus of the RT polypeptide. Also provided are compositions comprising the engineered RT or the recombinant RT protein and methods of using the engineered RT or the recombinant RT protein for performing reverse transcription reactions in a variety of applications. An exemplary DNA binding protein is DAT1. A full-length DNA binding protein, truncations or fragments thereof, 24 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    and/or peptide motifs with sequence specific binding can be used to engineer the RT contemplated by the present disclosure. [000126] The engineered RT polypeptides and/or the recombinant RT proteins disclosed herein were tested in 5’ single cell assay, 3’ single cell assay, and/or spatial platforms and were shown to significantly enhance the sensitivity of various single cell assays. (FIGs.20-33). [000127] FIG.20 and FIG.21 show that DAT1 in combination with reverse transcriptase MMLV, or MMLV variants thereof (e.g., SOLD 001 or SOLD 033 VDG), either fused at the N- terminus or the C-terminus of the RT improved the assay (GEX) sensitivity even at low sequencing depth. Improvement was observed with both gene expression, which was increased by up to ~37% and UMI, which was increased by up to 13% as captured at 20k rrpc. Gains were even more significant at higher read depth. In addition, the engineered RT molecules disclosed herein exhibited large change in differential gene expression in single cell assays. The engineered RT molecules comprising DAT1 or variants thereof disclosed herein picked-up up to about 5000 additional genes when compared to a non-DAT1 RT control (e.g., MMLV variant alone). FIGs.20-26. The engineered RT molecules disclosed herein also exhibited increase in median UMI counts per spot and median gene counts per spot in spatial assay. Moreover, the engineered RT molecules disclosed herein gave decrease in fraction of reads mapped to exons with gain in fraction mapped to introns. See e.g., FIG.21. A performance difference between FPLC and plate purified proteins was also demonstrated. [000128] The engineered RT polypeptides and/or the recombinant RT proteins disclosed herein were tested in spatial platforms using the 3’ workflow shown in FIG.27 under sandwich configuration conditions and using three different buffer formulations described herein. The engineered N-DAT1 RT variants disclosed herein enhanced the quality and the sensitivity of the spatial assay metrics while maintaining good spatial resolution when compared to a control RT lacking the DAT1 domain (also referred to as “control RT”). As used herein, a control RT, in the context of DAT1-RT fusion, refers to a non-DAT1 fusion RT comprising the same RT backbone as the N-DAT1 RT or the C-DAT1 RT. In some embodiments, the control RT is a MMLV variant (SEQ ID NOs: 1, 143, or 172) or MMLV variant L (SEQ ID NO: 145). [000129] As shown in FIGs.28A-F, the engineered RT polypeptides and/or the recombinant RT proteins disclosed herein maintained a good spatial resolution (e.g., an engineered RT variant 25 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    comprising DAT-1 at the N-terminus (N-DAT)). FIGs.28A-F show UMI heat maps showing globally detected gene expression in a spatial assay performed using a first slide configuration, on human tonsil tissue analyzed using NDAT1 RT variant (FIGs.28C-F) compared to a control RT variant (FIGs.28A-B) using three different buffer formulations. [000130] FIGs.29A-F show that the engineered RT polypeptides and/or the recombinant RT proteins disclosed herein (e.g., an engineered RT variant comprising DAT-1 at the N-terminus (N-DAT)) also significantly enhanced the quality and the sensitivity metrics of the spatial assays. FIGs.29A-F quantify the quality, sensitivity, and detection of gene expression metrics under the same six conditions shown in FIGs.28A-F. In particular, fraction reads in spots under tissue (FIG.29A) was substantially the same among the 6 conditions tested. However, the standard deviation of the sample comprising the N-DAT RT and RT reagent buffer (N-DAT_RTR) was smaller. The “fraction reads under tissue” refers to the ratio between reads in the area of direct interaction between the array and the tissue over total reads. The fraction reads under tissue showed the diffusion or transcript mislocalization that may have occurred during transcript release from the tissue or transcript capture onto the array. [000131] Reads mapped confidently to transcriptome (FIG.29B) were significantly increased in the samples containing the N-DAT_RT when compared to the five other conditions. The total genes detected (FIG.29C) appeared the same in all conditions tested, except in samples including the control RT variant and using RT reagent buffer (42B_RTR), which showed significantly reduced total genes detected. Analysis of GRch38 median UMI counts per spot (30k mapped spot-reads per spot) (FIG.29D) showed that median UMI counts per spot were significantly enhanced in all samples comprising an N-DAT RT variant when compared to the control RT samples regardless of the buffer used. Unexpectedly, samples comprising the N- DAT_RT variant and Buffer X.4 showed less variability. In addition, the GRch38 median genes per spot (30k mapped spot-reads per spot) (FIG.29E) analysis showed that median genes per spot were significantly enhanced in all samples comprising an N-DAT1 RT variant when compared to the control RT samples regardless of the buffer used. Assessment of individual gene expression also showed substantially similar results as globally detected gene expression. FIGs. 30A-F show UMI heat maps illustrating the gene expression of RGS3 in a human tonsil tissue. FIGs.31A-C show UMI heat maps illustrating the gene expression of KRT5 in a human tonsil 26 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    tissue. Additionally, spatial assays conducted with an N-DAT RT variant were successful using two different slide configurations: the first slide configuration (e.g., Visium standard definition (SD) spatial gene expression assay) or the second slide configuration (e.g., Visium high definition (HD) or next generation spatial gene expression assay), with a sample from Zebrafish (FIGs.32A-E and FIGs.33A-D). [000132] As such, the engineered RT polypeptides and/or the recombinant RT proteins disclosed herein improved the sensitivity of the spatial assay while maintaining good spatial resolution when compared to the control RT variant (FIGs.29D-E). Furthermore, in some metrics, such as reads mapping confidently to the genome, use of the N-DAT RT variant with RT Reagent was particularly superior.   [000133] Thus, the present disclosure demonstrates for the first time that incorporation of DAT1 or any molecule having substantially similar molecular function, in single and spatially assay increased transcript capture and single cell assay sensitivity. The DNA binding domain enhances the enzymatic activity of the engineered reverse transcriptase. For example, the addition of the DNA binding domain can enhance the template switching (TS) efficiency, higher end-to-end template jumping/switching, processivity efficiency, binding affinity, transcription efficiency, chemical tolerance, ability to yield mitochondrial unique molecular identifier (UMI) counts, ability to yield ribosomal unique molecular identifier (UMI) counts, shelf life, higher strand displacement, increased thermostability, improved thermoreactivity, and any combination thereof, for the engineered (i.e., recombinant) reverse transcriptase when compared to a control RT (SEQ ID NO: 1, 143, 145, or 172), WT MMLV, or known MMLV variants. [000134] TS efficiency: Small RNAs (<200 nucleotides) are for the most part non-coding regulatory elements and play a key role in gene expression. Small RNAs regulate gene expression in plants, animals, and many fungi—including several roles in development, proliferation, differentiation, immune reaction, apoptosis, tumorigenesis and adaptation to stress. Given their importance in regulation, miRNAs are candidates as biomarkers for several human diseases. Thus, developing accurate and reproducible ways to study these and other small RNAs is necessary to further decipher their biological consequences. [000135] The main sources of bias in a typical library preparation workflow are the enzymatic ligations that introduce 5′ and 3′ sequencing adaptors to single-stranded templates. Template 27 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    switching permits ligation-free incorporation of the 5′ adapter during reverse transcription. Template switching-based methods depend upon the natural tendency of MMLV-type reverse transcriptases to add nontemplated nucleotides at the 3′ end of the emerging cDNA strand. These nontemplated additions serve as an anchoring unit for annealing complementary nucleotides in a provided template switching oligonucleotide (TSO); upon reaching the cDNA-TSO cross- junction, the reverse transcriptase effectively switches templates, continuing cDNA synthesis out of the TSO sequence. By incorporating the 5′ adapter sequence into the TSO, and using polyadenylation to prime reverse transcription, ligation steps can be avoided altogether. For applications where the total RNA input is limited, such as single-cell RNA sequencing, template switching offers a critical advantage as it reduces the number of steps and sample loss during library preparation. Thus, the engineered reverse transcriptase described herein, exhibiting improved TS efficiency, is highly desirable. [000136] Higher end-to-end template jumping or switching: End-to-end template jumping or switching refers to the ability of a reverse transcriptase to template-switch from the 5’ end of one template to the 3’ end of another. Improved end-to-end template jumping or switching can result in an improved process efficiency. Thus, the engineered reverse transcriptase described herein, exhibiting improved or higher end-to-end template jumping or switching, is highly desirable. [000137] DNA binding affinity: To initiate reverse transcription, reverse transcriptases require a short DNA oligonucleotide called a primer to bind to its complementary sequences on the RNA template and serve as a starting point for synthesis of a new strand. Improved binding affinity results in a more efficient process, particularly when limited amounts of RNA are available. Thus, the engineered reverse transcriptase described herein, exhibiting improved DNA binding affinity, is highly desirable. [000138] Transcription efficiency: The RNA-to-cDNA conversion step in transcriptomics experiments is widely recognized as inefficient and variable. This issue is particularly significant for transcriptomics at the single cell level, which is preferable due to greater recognition of sample heterogeneity. Transcriptomics measurements almost invariably include a reverse transcription (RT) step, where RNA transcripts are used as templates to generate cDNA transcripts for quantification. This significantly complicates data interpretation as techniques are not directly measuring RNA transcript number, and results are therefore dependent on the 28 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    efficiency of the RNA to cDNA conversion. Thus, the engineered reverse transcriptase described herein, exhibiting improved transcription efficiency, is highly desirable. [000139] Chemical tolerance: Reverse transcriptases function in an environment that may include processing chemicals, such as cell fixation chemicals or processing reagents, which can negatively impact the function and activity of the enzyme. Thus, the engineered reverse transcriptase described herein, exhibiting improved chemical tolerance, is highly desirable. [000140] Ability to yield mitochondrial and/or ribosomal unique molecular identifier (UMI) counts: Unique molecular identifier (UMI) counting is a gene expression quantification scheme used in single-cell RNA-sequencing (scRNA-seq) analysis. Single-cell RNA-sequencing (scRNA-seq) technology provides transcriptome profiles of individual cells, enabling the dissection of the heterogeneity of different cell populations and tissues. The paucity of starting material for reverse transcription remains an inherent limitation of scRNA-seq protocols and contributes to the relatively low rate at which messenger RNA (mRNA) molecules in individual cells are converted to cDNA molecules that can be captured and sequenced. The miniscule quantity of transcripts captured from a single cell requires cDNA amplification for library construction; this inevitably results in large amplification bias. To mitigate this bias, some scRNA-seq protocols employ an additional step in which individual transcripts are barcoded with unique molecular identifiers (UMIs) before amplification, resulting in a more accurate quantification of the transcript count. UMIs incorporate a unique barcode onto each molecule within a given sample library. By incorporating individual barcodes on each original DNA fragment, variant alleles present in the original sample (true variants) can be distinguished from errors introduced during library preparation, target enrichment, or sequencing. Thus, the engineered reverse transcriptase described herein, exhibiting an improved ability to yield mitochondrial and/or ribosomal UMI counts, is highly desirable. [000141] Shelf life and/or stability: In another aspect of the disclosure, the engineered reverse transcriptase described herein, exhibit improved stability and/or shelf life. A longer period of stability, and/or shelf life, is desirable as it can result in more efficient processes. [000142] Higher strand displacement: Strand displacement is the process through which two strands with partial or full complementarity hybridize to each other, displacing one or more pre- hybridized strands in the process. Reverse transcriptase first transcribes a complementary strand 29 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    of DNA to make an RNA:DNA hybrid. Next, reverse transcriptase or RNase H degrades the RNA strand of the hybrid. The single-stranded DNA is then used as a template for synthesizing double-stranded DNA (cDNA). Thus, reverse transcriptase (RT) catalyzes the conversion of RNA into an integration-competent double-stranded DNA, with a variety of enzymatic activities that include the ability to displace a non-template strand concomitantly with polymerization. RT are capable of efficiently unwinding duplexes in the template during polymerization. This strand displacement synthesis activity by RT is required for the polymerization on the highly structured RNA and the removal of RNA fragments which cannot be cleaved by the enzymes’ RNase H activity. In addition, strand displacement synthesis on a DNA duplex is particularly important to complete the plus- and minus-strands by polymerizing on the long terminal repeats. As such, an RT with a higher strand displacement property is more efficient. Accordingly, the engineered reverse transcriptase described herein, exhibiting an improved strand displacement property, is highly desirable. [000143] Any of the engineered RT enzymes of the present disclosure, including without limitation any of the enzymes comprising the amino acid sequence and/or nucleic acid sequences shown in Table 1 or Table 2 could be analyzed in any suitable assay, including without limitation the assays described herein. Assays include without limitation 5’ gene expression analyses, with or without VDJ analysis, 3’ gene expression analysis, epigenetic analysis, or multiomic analyses. In non-limiting embodiments, experiments are carried out as found in the manufacturer’s instructions for the Chromium Single Cell 5’ Gene Expression Assay kit (10X Genomics); Chromium Single Cell 3’ Gene Expression Assay kit (10X Genomics), including any of multiomic extensions or applications. II. SPATIAL ANALYSIS METHODS [000144] Spatial analysis methodologies described herein can provide a vast amount of analyte and/or expression data for a variety of analytes within a biological sample at high spatial resolution, while retaining native spatial context. Spatial analysis methods can include, e.g., the use of a capture probe including a spatial barcode (e.g., a nucleic acid sequence that provides information as to the location or position of an analyte within a cell or a tissue sample (e.g., mammalian cell or a mammalian tissue sample) and a capture domain that is capable of binding to an analyte (e.g., a protein and/or a nucleic acid) produced by and/or present in a cell. Spatial 30 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    analysis methods and compositions can also include the use of a capture probe having a capture domain that captures an intermediate agent for indirect detection of an analyte. For example, the intermediate agent can include a nucleic acid sequence (e.g., a barcode) associated with the intermediate agent. Detection of the intermediate agent is therefore indicative of the analyte in the cell or tissue sample. [000145] Non-limiting aspects of spatial analysis methodologies and compositions are described in U.S. Patent Nos.11,447,807, 11,352,667, 11,168,350, 11,104,936, 11,008,608, 10,995,361, 10,913,975, 10,774,374, 10,724,078, 10,640,816, 10,494,662, 10,480,022, 10,364,457, 10,317,321, 10,059,990, 10,041,949, 10,030,261, 10,002,316, 9,879,313, 9,783,841, 9,727,810, 9,593,365, 8,951,726, 8,604,182, and 7,709,198; U.S. Patent Application Publication Nos. 2020/0239946, 2020/0080136, 2020/0277663, 2019/0330617, 2020/0256867, 2020/0224244, 2019/0085383, and 2013/0171621; PCT Publication Nos. WO2018/091676, WO2020/176788, WO2017/144338, and WO2016/057552; Non-patent literature references Rodriques et al., Science 363(6434):1463-1467, 2019; Lee et al., Nat. Protoc.10(3):442-458, 2015; Trejo et al., PLoS ONE 14(2):e0212031, 2019; Chen et al., Science 348(6233):aaa6090, 2015; Gao et al., BMC Biol.15:50, 2017; and Gupta et al., Nature Biotechnol.36:1197-1202, 2018; the Visium Spatial Gene Expression Reagent Kits User Guide (e.g., Rev F, dated January 2022); and/or the Visium Spatial Gene Expression Reagent Kits - Tissue Optimization User Guide (e.g., Rev E, dated February 2022), both of which are available at the 10x Genomics Support Documentation website, and can be used herein in any combination, and each of which is incorporated herein by reference in their entireties. Further non-limiting aspects of spatial analysis methodologies and compositions are described herein. [000146] Some general terminology that may be used in this disclosure can be found in Section (I)(b) of PCT Publication No. WO2020/176788 and/or U.S. Patent Application Publication No. 2020/0277663. Typically, a “barcode” is a label, or identifier, that conveys or is capable of conveying information (e.g., information about an analyte in a sample, a bead, and/or a capture probe). A barcode can be part of an analyte, or independent of an analyte. A barcode can be attached to an analyte. A particular barcode can be unique relative to other barcodes. For the purpose of this disclosure, an “analyte” can include any biological substance, structure, moiety, or component to be analyzed. The term “target” can similarly refer to an analyte of interest. 31 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    [000147] Analytes can be broadly classified into one of two groups: nucleic acid analytes, and non-nucleic acid analytes. Examples of non-nucleic acid analytes include, but are not limited to, lipids, carbohydrates, peptides, proteins, glycoproteins (N-linked or O-linked), lipoproteins, phosphoproteins, specific phosphorylated or acetylated variants of proteins, amidation variants of proteins, hydroxylation variants of proteins, methylation variants of proteins, ubiquitylation variants of proteins, sulfation variants of proteins, viral proteins (e.g., viral capsid, viral envelope, viral coat, viral accessory, viral glycoproteins, viral spike, etc.), extracellular and intracellular proteins, antibodies, and antigen binding fragments. In some embodiments, the analyte(s) can be localized to subcellular location(s), including, for example, organelles, e.g., mitochondria, Golgi apparatus, endoplasmic reticulum, chloroplasts, endocytic vesicles, exocytic vesicles, vacuoles, lysosomes, etc. In some embodiments, analyte(s) can be peptides or proteins, including without limitation antibodies and enzymes. Additional examples of analytes can be found in Section (I)(c) of PCT Publication No. WO2020/176788 and/or U.S. Patent Application Publication No.2020/0277663. In some embodiments, an analyte can be detected indirectly, such as through detection of an intermediate agent, for example, a ligation product or an analyte capture agent (e.g., an oligonucleotide-conjugated antibody), such as those described herein. [000148] A “biological sample” is typically obtained from the subject for analysis using any of a variety of techniques including, but not limited to, biopsy, surgery, and laser capture microscopy (LCM), and generally includes cells and/or other biological material from the subject. In some embodiments, the biological sample is a tissue sample. In some embodiments, the biological sample (e.g., tissue sample) is a tissue microarray (TMA). A tissue microarray contains multiple representative tissue samples – which can be from different tissues or organisms – assembled on a single histologic slide. The TMA can therefore allow for high throughput analysis of multiple specimens at the same time. Tissue microarrays are paraffin blocks produced by extracting cylindrical tissue cores from different paraffin donor blocks and re-embedding these into a single recipient (microarray) block at defined array coordinates. [000149] The biological sample as used herein can be any suitable biological sample described herein or known in the art. In some embodiments, the biological sample is a tissue. In some embodiments, the tissue sample is a solid tissue sample. In some embodiments, the biological sample is a tissue section. In some embodiments, the tissue is flash-frozen and sectioned. Any 32 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    suitable method described herein or known in the art can be used to flash-freeze and section the tissue sample. In some embodiments, the biological sample, e.g., the tissue, is flash-frozen using liquid nitrogen before sectioning. In some embodiments, the biological sample, e.g., a tissue sample, is flash-frozen using nitrogen (e.g., liquid nitrogen), isopentane, or hexane. [000150] In some embodiments, the biological sample, e.g., the tissue, is embedded in a matrix e.g., optimal cutting temperature (OCT) compound to facilitate sectioning. OCT compound is a formulation of clear, water-soluble glycols and resins, providing a solid matrix to encapsulate biological (e.g., tissue) specimens. In some embodiments, the sectioning is performed using cryosectioning. In some embodiments, the methods further comprise a thawing step, after the cryosectioning. [000151] The biological sample can be from a mammal. In some instances, the biological sample is from a human, mouse, or rat. In addition to the subjects described above, the biological sample can be obtained from non-mammalian organisms (e.g., a plants, an insect, an arachnid, a nematode (e.g., Caenorhabditis elegans), a fungi, an amphibian, or a fish (e.g., zebrafish)). A biological sample can be obtained from a prokaryote such as a bacterium, e.g., Escherichia coli, Staphylococci or Mycoplasma pneumoniae; an archaea; a virus such as Hepatitis C virus or human immunodeficiency virus; or a viroid. A biological sample can be obtained from a eukaryote, such as a patient derived organoid (PDO) or patient derived xenograft (PDX). The biological sample can include organoids, a miniaturized and simplified version of an organ produced in vitro in three dimensions that shows realistic micro-anatomy. Organoids can be generated from one or more cells from a tissue, embryonic stem cells, and/or induced pluripotent stem cells, which can self-organize in three-dimensional culture owing to their self-renewal and differentiation capacities. In some embodiments, an organoid is a cerebral organoid, an intestinal organoid, a stomach organoid, a lingual organoid, a thyroid organoid, a thymic organoid, a testicular organoid, a hepatic organoid, a pancreatic organoid, an epithelial organoid, a lung organoid, a kidney organoid, a gastruloid, a cardiac organoid, or a retinal organoid. Subjects from which biological samples can be obtained can be healthy or asymptomatic individuals, individuals that have or are suspected of having a disease (e.g., cancer) or a pre-disposition to a disease, and/or individuals that are in need of therapy or suspected of needing therapy. 33 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    [000152] Biological samples can be derived from a homogeneous culture or population of the subjects or organisms mentioned herein or alternatively from a collection of several different organisms, for example, in a community or ecosystem. [000153] Biological samples can include one or more diseased cells. A diseased cell can have altered metabolic properties, gene expression, protein expression, and/or morphologic features. Examples of diseases include inflammatory disorders, metabolic disorders, nervous system disorders, and cancer. Cancer cells can be derived from solid tumors, hematological malignancies, cell lines, or obtained as circulating tumor cells. [000154] In some embodiments, the biological sample, e.g., the tissue sample, is fixed in a fixative including alcohol, for example methanol. In some embodiments, instead of methanol, acetone, or an acetone-methanol mixture can be used. In some embodiments, the fixation is performed after sectioning. In some instances, the biological sample is not fixed with paraformaldehyde (PFA). In some instances, when the biological sample is fixed with a fixative including an alcohol (e.g., methanol or acetone-methanol mixture), it is not decrosslinked afterward. In some preferred embodiments, the biological sample is fixed with a fixative including an alcohol (e.g., methanol or an acetone-methanol mixture) after freezing and/or sectioning. In some instances, the biological sample is flash-frozen, and then the biological sample is sectioned and fixed (e.g., using methanol, acetone, or an acetone-methanol mixture). In some instances when methanol, acetone, or an acetone-methanol mixture is used to fix the biological sample, the sample is not decrosslinked at a later step. In instances when the biological sample is frozen (e.g., flash frozen using liquid nitrogen and embedded in OCT) followed by sectioning and alcohol (e.g., methanol, acetone-methanol) fixation or acetone fixation, the biological sample is referred to as “fresh frozen”. In some embodiments, fixation of the biological sample e.g., using acetone and/or alcohol (e.g., methanol, acetone-methanol) is performed while the sample is mounted on a substrate (e.g., glass slide, such as a positively charged glass slide). [000155] In some embodiments, the biological sample, e.g., the tissue sample, is fixed e.g., immediately after being harvested from a subject. In such embodiments, the fixative is preferably an aldehyde fixative, such as paraformaldehyde (PFA) or formalin. In some embodiments, the fixative induces crosslinks within the biological sample. In some embodiments, after fixing e.g., 34 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    by formalin or PFA, the biological sample is dehydrated via sucrose gradient. In some instances, the fixed biological sample is treated with a sucrose gradient and then embedded in a matrix e.g., OCT compound. In some instances, the fixed biological sample is not treated with a sucrose gradient, but rather is embedded in a matrix e.g., OCT compound after fixation. In some embodiments when a fixed frozen tissue sample is treated with a sucrose gradient, it can be rehydrated with an ethanol gradient. In some embodiments, the PFA or formalin fixed biological sample, which can be optionally dehydrated via sucrose gradient and/or embedded in OCT compound, is then frozen e.g., for storage or shipment. In such instances, the biological sample is referred to as “fixed frozen”. In preferred embodiments, a fixed frozen biological sample is not treated with methanol. In preferred embodiments, a fixed frozen biological sample is not paraffin embedded. Thus, in preferred embodiments, a fixed frozen biological sample is not deparaffinized. In some embodiments, a fixed frozen biological sample is rehydrated in an ethanol gradient. [000156] In some instances, the biological sample (e.g., a fixed frozen tissue sample) is treated with a citrate buffer. Citrate buffer can be used for antigen retrieval to decrosslink antigens and fixation medium in the biological sample. Thus, any suitable decrosslinking agent can be used in addition to or alternatively to citrate buffer. In some embodiments, for example, the biological sample (e.g., a fixed frozen tissue sample) is decrosslinked with TE buffer. [000157] In any of the foregoing, the biological sample can further be stained, imaged, and/or destained. For example, in some embodiments, a fresh frozen tissue sample or fixed frozen tissue sample is stained (e.g., via eosin and/or hematoxylin), imaged, destained (e.g., via HCl), or a combination thereof. In some embodiments, when a fresh frozen tissue sample is fixed in methanol, it is treated with isopropanol prior to being stained (e.g., via eosin and/or hematoxylin), imaged, destained (e.g., via HCl), or a combination thereof. In some embodiments when a fixed frozen tissue sample is treated with a sucrose gradient, it can be rehydrated with an ethanol gradient before being stained, (e.g., via eosin and/or hematoxylin), imaged, destained (e.g., via HCl), decrosslinked (e.g., via TE buffer or citrate buffer), or a combination thereof. In some embodiments, the biological sample can undergo further fixation (e.g., while mounted on a substrate), stained, imaged, and/or destained. For example, a fixed frozen biological sample may 35 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    be subject to an additional fixing step (e.g., using PFA) before optional ethanol rehydration, staining, imaging, and/or destaining. [000158] In any of the foregoing, the biological sample can be fixed using PAXgene®. For example, the biological sample can be fixed using PAXgene® in addition, or alternatively to, a fixative disclosed herein or known in the art (e.g., alcohol, acetone, acetone-alcohol, formalin, paraformaldehyde). PAXgene® is a non-cross-linking mixture of different alcohols, acid and a soluble organic compound that preserves morphology and bio-molecules. It is a two-reagent fixative system in which tissue is firstly fixed in a solution containing methanol and acetic acid then stabilized in a solution containing ethanol. See, Ergin B. et al., J Proteome Res.2010 Oct 1;9(10):5188-96; Kap M. et al., PLoS One.; 6(11):e27704 (2011); and Mathieson W. et al., Am J Clin Pathol.; 146(1):25-40 (2016), each of which are hereby incorporated by reference in their entirety, for a description and evaluation of PAXgene® for tissue fixation. Thus, in some embodiments, when the biological sample, e.g., the tissue sample, is fixed in a fixative including alcohol, the fixative is PAXgene®. In some embodiments, a fresh frozen tissue sample is fixed with PAXgene®. In some embodiments, a fixed frozen tissue sample is fixed with PAXgene®. [000159] In some embodiments, the biological sample, e.g., the tissue sample is fixed, for example in methanol, acetone, acetone-methanol, PFA, PAXgene® or is formalin-fixed and paraffin-embedded (FFPE). In some embodiments, the biological sample comprises intact cells. In some embodiments, the biological sample is a cell pellet, e.g., a fixed cell pellet, e.g., an FFPE cell pellet. FFPE samples are used in some instances in the RTL methods disclosed herein. A limitation of direct RNA capture for fixed samples is that the RNA integrity of fixed (e.g., FFPE) samples can be lower than a fresh sample, thereby making it more difficult to capture RNA directly, e.g., by capture of a common sequence such as a poly(A) tail of an mRNA molecule. However, by utilizing RTL probes that hybridize to RNA target sequences in the transcriptome, one can avoid a requirement for RNA analytes to have both a poly(A) tail and target sequences intact. Accordingly, RTL probes can be utilized to beneficially improve capture and spatial analysis of fixed samples. The biological sample, e.g., tissue sample, can be stained, and imaged prior, during, and/or after each step of the methods described herein. Any of the methods described herein or known in the art can be used to stain and/or image the biological sample. In some embodiments, the imaging occurs prior to destaining the sample. In some embodiments, 36 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    the biological sample is stained using an H&E staining method. In some embodiments, the tissue sample is stained and imaged for about 10 minutes to about 2 hours (or any of the subranges of this range described herein). Additional time may be needed for staining and imaging of different types of biological samples. [000160] The tissue sample can be obtained from any suitable location in a tissue or organ of a subject, e.g., a human subject. In some instances, the sample is a mouse sample. In some instances, the sample is a human sample. In some embodiments, the sample can be derived from skin, brain, breast, lung, liver, kidney, prostate, tonsil, thymus, testes, bone, lymph node, ovary, eye, heart, or spleen. In some instances, the sample is a human or mouse breast tissue sample. In some instances, the sample is a human or mouse brain tissue sample. In some instances, the sample is a human or mouse lung tissue sample. In some instances, the sample is a human or mouse tonsil tissue sample. In some instances, the sample is a human or mouse liver tissue sample. In some instances, the sample is a human or mouse bone, skin, kidney, thymus, testes, or prostate tissue sample. In some embodiments, the tissue sample is derived from normal or diseased tissue. In some embodiments, the sample is an embryo sample. The embryo sample can be a non-human embryo sample. In some instances, the sample is a mouse embryo sample. [000161] Non-limiting examples of stains include histological stains (e.g., hematoxylin and/or eosin) and immunological stains (e.g., fluorescent stains). The biological sample can be stained using Can-Grunwald, Giemsa, hematoxylin and eosin (H&E), Jenner’s, Leishman, Masson’s trichrome, Papanicolaou, Romanowsky, silver, Sudan, Wright’s, and/or Periodic Acid Schiff (PAS) staining techniques. In some instances, PAS staining is performed after formalin or acetone fixation. In some embodiments, a biological sample (e.g., a fixed and/or stained biological sample) can be imaged. Biological samples are also described in Section (I)(d) of PCT Publication No. WO2020/176788 and/or U.S. Patent Application Publication No.2020/0277663. [000162] The following embodiments can be used with any of the methods described herein. In some embodiments, the biological sample is imaged. In some embodiments, the biological sample is visualized or imaged using bright field microscopy. In some embodiments, the biological sample is visualized or imaged using fluorescence microscopy. Additional methods of visualization and imaging are known in the art. Non-limiting examples of visualization and imaging include expansion microscopy, bright field microscopy, dark field microscopy, phase 37 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    contrast microscopy, electron microscopy, fluorescence microscopy, reflection microscopy, interference microscopy and confocal microscopy. In some embodiments, the sample is stained and imaged prior to adding the primer to the biological sample. [000163] In some embodiments, the method includes staining the biological sample. In some embodiments, the staining includes the use of hematoxylin and eosin. In some embodiments, a biological sample can be stained using any number of biological stains, including but not limited to, acridine orange, Bismarck brown, carmine, coomassie blue, cresyl violet, DAPI, eosin, ethidium bromide, acid fuchsine, hematoxylin, Hoechst stains, iodine, methyl green, methylene blue, neutral red, Nile blue, Nile red, osmium tetroxide, propidium iodide, rhodamine, or safranin. In some instances, the biological sample can be stained using known staining techniques, including Can-Grunwald, Giemsa, hematoxylin and eosin (H&E), Jenner’s, Leishman, Masson’s trichrome, Papanicolaou, Romanowsky, silver, Sudan, Wright’s, and/or Periodic Acid Schiff (PAS) staining techniques. PAS staining is typically performed after formalin or acetone fixation. [000164] In some embodiments, the staining includes the use of a detectable label selected from the group consisting of a radioisotope, a fluorophore, a chemiluminescent compound, a bioluminescent compound, or a combination thereof. [000165] In some embodiments, a biological sample is permeabilized with one or more permeabilization reagents. For example, permeabilization of a biological sample can facilitate analyte capture. Exemplary permeabilization agents and conditions are described in Section (I)(d)(ii)(13) or the Exemplary Embodiments Section of PCT Publication No. WO2020/176788 and/or U.S. Patent Application Publication No.2020/0277663. Briefly, in any of the methods described herein, the method includes a step of permeabilizing the biological sample. For example, the biological sample can be permeabilized to facilitate transfer of the extension products to the capture probes on the array. In some embodiments, the permeabilizing includes the use of an organic solvent (e.g., acetone, ethanol, and methanol), a detergent (e.g., saponin, Triton X-100™, Tween-20™, or sodium dodecyl sulfate (SDS)), an enzyme (an endopeptidase, an exopeptidase, a protease), or combinations thereof. In some embodiments, the permeabilizing includes the use of an endopeptidase, a protease, SDS, polyethylene glycol tert-octylphenyl ether, polysorbate 80, and polysorbate 20, N-lauroylsarcosine sodium salt solution, saponin, 38 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    Triton X-100™, Tween-20™, or combinations thereof. In some embodiments, the endopeptidase is pepsin. In some embodiments, the endopeptidase is Proteinase K. Additional methods for sample permeabilization are described, for example, in Jamur et al., Method Mol. Biol.588:63- 66, 2010, the entire contents of which are incorporated herein by reference. [000166] Array-based spatial analysis methods involve the transfer of one or more analytes from a biological sample to an array of features on a substrate, where each feature is associated with a unique spatial location on the array. Subsequent analysis of the transferred analytes includes determining the identity of the analytes and the spatial location of the analytes within the biological sample. The spatial location of an analyte within the biological sample is determined based on the feature to which the analyte is bound (e.g., directly or indirectly) on the array, and the feature’s relative spatial location within the array. [000167] A “capture probe” refers to any molecule capable of capturing (directly or indirectly) and/or labelling an analyte (e.g., an analyte of interest) in a biological sample. In some embodiments, the capture probe is a nucleic acid or a polypeptide. In some embodiments, the capture probe includes a barcode (e.g., a spatial barcode and/or a unique molecular identifier (UMI)) and a capture domain). In some instances, the capture probe includes a homopolymer sequence, such as a poly(T) sequence. In some embodiments, a capture probe can include a cleavage domain and/or a functional domain (e.g., a primer-binding site, such as for next- generation sequencing (NGS)). See, e.g., Section (II)(b) (e.g., subsections (i)-(vi)) of PCT Publication No. WO2020/176788 and/or U.S. Patent Application Publication No.2020/0277663. Generation of capture probes can be achieved by any appropriate method, including those described in Section (II)(d)(ii) of PCT Publication No. WO2020/176788 and/or U.S. Patent Application Publication No.2020/0277663. [000168] In some embodiments, the biological sample is mounted on a first substrate and the substrate comprising the array of capture probes is a second substrate. During this process, one or more analytes or analyte derivatives (e.g., intermediate agents; e.g., ligation products) are released from the biological sample and migrate to the second substrate comprising an array of capture probes. In some embodiments, the release and migration of the analytes or analyte derivatives to the second substrate comprising the array of capture probes occurs in a manner that preserves the original spatial context of the analytes in the biological sample. This method 39 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    can be referred to as a sandwiching process, which is described e.g., in U.S. Patent Application Pub. No.2021/0189475 and PCT Pub. Nos. WO 2021/252747 A1, WO 2022/061152 A2, and WO 2022/140028 A1. [000169] FIG.1A shows an exemplary sandwiching process 100 where a first substrate (e.g., slide 103), including a biological sample 102, and a second substrate (e.g., array slide 104 including an array having spatially barcoded capture probes 106) are brought into proximity with one another. As shown in FIG.1A a liquid reagent drop (e.g., permeabilization solution 105) is introduced on the second substrate in proximity to the capture probes 106 and in between the biological sample 102 and the second substrate (e.g., slide 104 including an array having spatially barcoded capture probes 106). The permeabilization solution 105 may release analytes or analyte derivatives (e.g., intermediate agents; e.g., ligation products) that can be captured by the capture probes of the array 106. [000170] During the exemplary sandwiching process, the first substrate is aligned with the second substrate, such that at least a portion of the biological sample is aligned with at least a portion of the capture probes (e.g., aligned in a sandwich configuration). As shown, the second substrate (e.g., array slide 104) is in an inferior position to the first substrate (e.g., slide 103). In some embodiments, the first substrate (e.g., slide 103) may be positioned superior to the second substrate (e.g., slide 104). A reagent medium 105 within a gap between the first substrate (e.g., slide 103) and the second substrate (e.g., slide 104) creates a liquid interface between the two substrates. The reagent medium may be a permeabilization solution which permeabilizes and/or digests the biological sample 102. In some embodiments wherein the biological sample 102 has been pre-permeabilized, the reagent medium is not a permeabilization solution. In some embodiments, analytes (e.g., mRNA transcripts) and/or analyte derivatives (e.g., intermediate agents; e.g., ligation products) of the biological sample 102 may release from the biological sample, and actively or passively migrate (e.g., diffuse) across the gap toward the capture probes on the array 106. Alternatively, in certain embodiments, migration of the analyte or analyte derivative (e.g., intermediate agent; e.g., ligation product) from the biological sample is performed actively (e.g., electrophoretic, by applying an electric field to promote migration). Exemplary methods of electrophoretic migration are described in WO 2020/176788, and US. Patent Application Pub. No.2021/0189475, each of which is hereby incorporated by reference. 40 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    [000171] As further shown, one or more spacers 110 may be positioned between the first substrate (e.g., slide 103) and the second substrate (e.g., array slide 104 including spatially barcoded capture probes 106). The one or more spacers 110 may be configured to maintain a separation distance between the first substrate and the second substrate. While the one or more spacers 110 is shown as disposed on the second substrate, the spacer may additionally or alternatively be disposed on the first substrate. [000172] In some embodiments, the one or more spacers 110 is configured to maintain a separation distance between first and second substrates that is between about 2 microns and 1 mm (e.g., between about 2 microns and 800 microns, between about 2 microns and 700 microns, between about 2 microns and 600 microns, between about 2 microns and 500 microns, between about 2 microns and 400 microns, between about 2 microns and 300 microns, between about 2 microns and 200 microns, between about 2 microns and 100 microns, between about 2 microns and 25 microns, or between about 2 microns and 10 microns), measured in a direction orthogonal to the surface of first substrate that supports the biological sample. In some instances, the separation distance is about 2, 3, 4, 5, 6, 7, 8, 9, 10, 11, 12, 13, 14, 15, 16, 17, 18, 19, 20, 21, 22, 23, 24, or 25 microns. In some embodiments, the separation distance is less than 50 microns. In some embodiments, the separation distance is less than 25 microns. In some embodiments, the separation distance is less than 20 microns. The separation distance may include a distance of at least 2 µm. [000173] FIG.1B shows a fully formed sandwich configuration 125 creating a chamber 150 formed from the one or more spacers 110, the first substrate (e.g., the slide 103), and the second substrate (e.g., the slide 104 including an array 106 having spatially barcoded capture probes) in accordance with some example implementations. In the example of FIG.1B, the liquid reagent (e.g., the permeabilization solution 105) fills the volume of the chamber 150 and may create a permeabilization buffer that allows analytes (e.g., mRNA transcripts and/or other molecules) or analyte derivatives (e.g., intermediate agents; e.g., ligation products) to diffuse from the biological sample 102 toward the capture probes of the second substrate (e.g., slide 104). [000174] In some aspects, flow of the permeabilization buffer may deflect transcripts and/or molecules from the biological sample 102 and may affect diffusive transfer of analytes or analyte derivatives (e.g., intermediate agents; e.g., ligation products) for spatial analysis. A partially or 41 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    fully sealed chamber 150 resulting from the one or more spacers 110, the first substrate, and the second substrate may reduce or prevent flow from undesirable convective movement of transcripts and/or molecules over the diffusive transfer from the biological sample 102 to the capture probes. [000175] The sandwiching process methods described above can be implemented using a variety of hardware components. For example, the sandwiching process methods can be implemented using a sample holder (also referred to herein as a support device, a sample handling apparatus, and an array alignment device). Further details on support devices, sample holders, sample handling apparatuses, or systems for implementing a sandwiching process are described in, e.g., US. Patent Application Pub. No.2021/0189475, and PCT Publ. No. WO 2022/061152 A2, each of which are incorporated by reference in their entirety. [000176] In some embodiments of a sample holder, the sample holder can include a first member including a first retaining mechanism configured to retain a first substrate comprising a biological sample. The first retaining mechanism can be configured to retain the first substrate disposed in a first plane. The sample holder can further include a second member including a second retaining mechanism configured to retain a second substrate disposed in a second plane. The sample holder can further include an alignment mechanism connected to one or both of the first member and the second member. The alignment mechanism can be configured to align the first and second members along the first plane and/or the second plane such that the sample contacts at least a portion of the reagent medium when the first and second members are aligned and within a threshold distance along an axis orthogonal to the second plane. The adjustment mechanism may be configured to move the second member along the axis orthogonal to the second plane and/or move the first member along an axis orthogonal to the first plane. [000177] In some embodiments, the adjustment mechanism includes a linear actuator. In some embodiments, the linear actuator is configured to move the second member along an axis orthogonal to the plane of the first member and/or the second member. In some embodiments, the linear actuator is configured to move the first member along an axis orthogonal to the plane of the first member and/or the second member. In some embodiments, the linear actuator is configured to move the first member, the second member, or both the first member and the second member at a velocity of at least 0.1 mm/sec. In some embodiments, the linear actuator is 42 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    configured to move the first member, the second member, or both the first member and the second member with an amount of force of at least 0.1 lbs. [000178] FIG.2A is a perspective view of an example sample handling apparatus 200 in a closed position in accordance with some example implementations. As shown, the sample handling apparatus 200 includes a first member 204, a second member 210, optionally an image capture device 220, a first substrate 206, optionally a hinge 215, and optionally a mirror 216. The hinge 215 may be configured to allow the first member 204 to be positioned in an open or closed configuration by opening and/or closing the first member 204 in a clamshell manner along the hinge 215. [000179] FIG.2B is a perspective view of the example sample handling apparatus 200 in an open position in accordance with some example implementations. As shown, the sample handling apparatus 200 includes one or more first retaining mechanisms 208 configured to retain one or more first substrates 206. In the example of FIG.2B, the first member 204 is configured to retain two first substrates 206, however the first member 204 may be configured to retain more or fewer first substrates 206. [000180] In some aspects, when the sample handling apparatus 200 is in an open position (e.g., in FIG.2B), the first substrate 206 and/or the second substrate 212 may be loaded and positioned within the sample handling apparatus 200 such as within the first member 204 and the second member 210, respectively. As noted, the hinge 215 may allow the first member 204 to close over the second member 210 and form a sandwich configuration. [000181] In some aspects, after the first member 204 closes over the second member 210, an adjustment mechanism of the sample handling apparatus 200 may actuate the first member 204 and/or the second member 210 to form the sandwich configuration for the permeabilization step (e.g., bringing the first substrate 206 and the second substrate 212 closer to each other and within a threshold distance for the sandwich configuration). The adjustment mechanism may be configured to control a speed, an angle, a force, or the like of the sandwich configuration. [000182] In some embodiments, the biological sample (e.g., sample 102 from FIG.1A) may be aligned within the first member 204 (e.g., via the first retaining mechanism 208) prior to closing the first member 204 such that a desired region of interest of the sample is aligned with the 43 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    barcoded array of the second substrate (e.g., the slide 104 from FIG.1A), e.g., when the first and second substrates are aligned in the sandwich configuration. Such alignment may be accomplished manually (e.g., by a user) or automatically (e.g., via an automated alignment mechanism). After or before alignment, spacers may be applied to the first substrate 206 and/or the second substrate 212 to maintain a minimum spacing between the first substrate 206 and the second substrate 212 during sandwiching. In some aspects, the permeabilization solution (e.g., permeabilization solution 305) may be applied to the first substrate 206 and/or the second substrate 212. The first member 204 may then close over the second member 210 and form the sandwich configuration. Analytes or analyte derivatives (e.g., intermediate agents; e.g., ligation products) may be captured by the capture probes of the array and may be processed for spatial analysis. [000183] In some embodiments, during the permeabilization step, the image capture device 220 may capture images of the overlap area between the biological sample and the capture probes on the array 106. If more than one first substrates 206 and/or second substrates 212 are present within the sample handling apparatus 200, the image capture device 220 may be configured to capture one or more images of one or more overlap areas. [000184] Provided herein are methods for delivering a fluid to a biological sample disposed on an area of a first substrate and an array disposed on a second substrate. FIGs.3A-3C depict a side view and a top view of an exemplary angled closure workflow 300 for sandwiching a first substrate (e.g., slide 303) having a biological sample 302 and a second substrate (e.g., slide 304 having capture probes 306) in accordance with some exemplary implementations. [000185] FIG.3A depicts the first substrate (e.g., the slide 303 including a biological sample 302) angled over (superior to) the second substrate (e.g., slide 304). As shown, reagent medium (e.g., permeabilization solution) 305 is located on the spacer 310 toward the right-hand side of the side view in FIG.3A. While FIG.3A depicts the reagent medium on the right hand side of side view, it should be understood that such depiction is not meant to be limiting as to the location of the reagent medium on the spacer. [000186] FIG.3B shows that as the first substrate lowers, and/or as the second substrate rises, the dropped side of the first substrate (e.g., a side of the slide 303 angled toward the second substrate) may contact the reagent medium 305. The dropped side of the first substrate may urge 44 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    the reagent medium 305 toward the opposite direction (e.g., towards an opposite side of the spacer 310, towards an opposite side of the first substrate relative to the dropped side). For example, in the side view of FIG.3B the reagent medium 305 may be urged from right to left as the sandwich is formed. [000187] In some embodiments, the first substrate and/or the second substrate are further moved to achieve an approximately parallel arrangement of the first substrate and the second substrate. [000188] FIG.3C depicts a full closure of the sandwich between the first substrate and the second substrate with the spacer 310 contacting both the first substrate and the second substrate and maintaining a separation distance and optionally the approximately parallel arrangement between the two substrates. As shown in the top view of FIG.3C, the spacer 310 fully encloses and surrounds the biological sample 302 and the capture probes 306, and the spacer 310 form the sides of chamber 350 which holds a volume of the reagent medium 305. [000189] While FIG.3C depicts the first substrate (e.g., the slide 303 including biological sample 302) angled over (superior to) the second substrate (e.g., slide 304) and the second substrate comprising the spacer 310, it should be understood that an exemplary angled closure workflow can include the second substrate angled over (superior to) the first substrate and the first substrate comprising the spacer 310. [000190] It may be desirable that the reagent medium be free from air bubbles between the substrates to facilitate transfer of target analytes with spatial information. Additionally, air bubbles present between the substrates may obscure at least a portion of an image capture of a desired region of interest. Accordingly, it may be desirable to ensure or encourage suppression and/or elimination of air bubbles between the two substrates (e.g., slide 303 and slide 304) during a permeabilization step (e.g., step 104). In some aspects, it may be possible to reduce or eliminate bubble formation between the substrates using a variety of filling methods and/or closing methods. In some instances, the first substrate and the second substrate are arranged in an angled sandwich assembly as described herein. For example, during the sandwiching of the two substrates (e.g., the slide 303 and the slide 304), an angled closure workflow may be used to suppress or eliminate bubble formation. 45 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    [000191] FIG.4A is a side view of the angled closure workflow 400 in accordance with some exemplary implementations. FIG.4B is a top view of the angled closure workflow 400 in accordance with some exemplary implementations. As shown at 405, reagent medium 401 is positioned to the side of the substrate 402. [000192] At step 410, the dropped side of the angled substrate 406 contacts the reagent medium 401 first. The contact of the substrate 406 with the reagent medium 401 may form a linear or low curvature flow front that fills uniformly with the slides closed. [000193] At step 415, the substrate 406 is further lowered toward the substrate 402 (or the substrate 402 is raised up toward the substrate 406) and the dropped side of the substrate 406 may contact and may urge the reagent medium toward the side opposite the dropped side and creating a linear or low curvature flow front that may prevent or reduce bubble trapping between the substrates. [000194] At step 420, the reagent medium 401 fills the gap between the substrate 406 and the substrate 402. The linear flow front of the liquid reagent may form by squeezing the 401 volume along the contact side of the substrate 402 and/or the substrate 406. Additionally, capillary flow may also contribute to filling the gap area. [000195] In some embodiments, the reagent medium (e.g., 105 in FIG 1A) comprises a permeabilization agent. In some embodiments, following initial contact between the biological sample and a permeabilization agent, the permeabilization agent can be removed from contact with the biological sample (e.g., by opening sample holder). Suitable agents for this purpose include, but are not limited to, organic solvents (e.g., acetone, ethanol, and methanol), cross- linking agents (e.g., paraformaldehyde), detergents (e.g., saponin, Triton X-100™, Tween-20™, or sodium dodecyl sulfate (SDS)), and enzymes (e.g., trypsin, proteases (e.g., proteinase K). In some embodiments, the detergent is an anionic detergent (e.g., SDS or N-lauroylsarcosine sodium salt solution). [000196] In some embodiments, the reagent medium comprises a lysis reagent. Lysis solutions can include ionic surfactants such as, for example, sarkosyl and sodium dodecyl sulfate (SDS). More generally, chemical lysis agents can include, without limitation, organic solvents, chelating agents, detergents, surfactants, and chaotropic agents. In some embodiments, the reagent 46 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    medium comprises a protease. Exemplary proteases include, e.g., pepsin, trypsin, pepsin, elastase, and proteinase K. In some embodiments, the reagent medium comprises a nuclease. In some embodiments, the nuclease comprises an RNase. In some embodiments, the Rnase is selected from Rnase A, Rnase C, Rnase H, and Rnase I. In some embodiments, the reagent medium comprises one or more of sodium dodecyl sulfate (SDS) or a sodium salt thereof, proteinase K, pepsin, N-lauroylsarcosine, and RNAse. [000197] In some embodiments, the reagent medium comprises polyethylene glycol (PEG). In some embodiments, the PEG is from about PEG 2K to about PEG 16K. In some embodiments, the PEG is PEG 2K, 3K, 4K, 5K, 6K, 7K, 8K, 9K, 10K, 11K, 12K, 13K, 14K, 15K, or 16K. In some embodiments, the PEG is present at a concentration from about 2% to 25%, from about 4% to about 23%, from about 6% to about 21%, or from about 8% to about 20% (v/v). [000198] In certain embodiments a dried permeabilization reagent is applied or formed as a layer on the first substrate or the second substrate or both prior to contacting the biological sample and the array. For example, a permeabilization reagent can be deposited in solution on the first substrate or the second substrate or both and then dried. [000199] In some instances, the aligned portions of the biological sample and the array are in contact with the reagent medium for about 1 minute, about 5 minutes, about 10 minutes, about 12 minutes, about 15 minutes, about 18 minutes, about 20 minutes, about 25 minutes, about 30 minutes, about 36 minutes, about 45 minutes, or about an hour. In some instances, the aligned portions of the biological sample and the array are in contact with the reagent medium for about 1-60 minutes. [000200] In some instances, the device is configured to control a temperature of the first and second substrates. In some embodiments, the temperature of the first and second members is lowered to a first temperature that is below room temperature. [000201] There are at least two methods to associate a spatial barcode with one or more neighboring cells, such that the spatial barcode identifies the one or more cells, and/or contents of the one or more cells, as associated with a particular spatial location. One method is to promote analytes or analyte proxies (e.g., intermediate agents) out of a cell and towards a spatially-barcoded array (e.g., including spatially-barcoded capture probes). Another method is 47 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    to cleave spatially-barcoded capture probes from an array and promote the spatially-barcoded capture probes towards and/or into or onto the biological sample. [000202] In some cases, capture probes may be configured to prime, replicate, and consequently yield optionally barcoded extension products from a template (e.g., a DNA or RNA template, such as an analyte or an intermediate agent (e.g., a ligation product or an analyte capture agent), or a portion thereof), or derivatives thereof (see, e.g., Section (II)(b)(vii) of PCT Publication No. WO2020/176788 and/or U.S. Patent Application Publication No.2020/0277663 regarding extended capture probes). In some cases, capture probes may be configured to form ligation products with a template (e.g., a DNA or RNA template, such as an analyte or an intermediate agent, or portion thereof), thereby creating ligations products that serve as proxies for the template. [000203] As used herein, an “extended capture probe” refers to a capture probe having additional nucleotides added to the terminus (e.g., 3’ or 5’ end) of the capture probe thereby extending the overall length of the capture probe. For example, an “extended 3’ end” indicates additional nucleotides were added to the most 3’ nucleotide of the capture probe to extend the length of the capture probe, for example, by polymerization reactions used to extend nucleic acid molecules including templated polymerization catalyzed by a polymerase (e.g., a DNA polymerase or a reverse transcriptase). In some embodiments, extending the capture probe includes adding to a 3’ end of a capture probe a nucleic acid sequence that is complementary to a nucleic acid sequence of an analyte or intermediate agent specifically bound to the capture domain of the capture probe. In some embodiments, the capture probe is extended by a reverse transcriptase. In some embodiments, the capture probe is extended using one or more DNA polymerases. In some embodiments, the extended capture probes include the sequence of the capture domain and the sequence of the spatial barcode of the capture probe. [000204] In some embodiments, extended capture probes are amplified (e.g., in bulk solution or on the array) to yield quantities that are sufficient for downstream analysis, e.g., sequencing. In some embodiments, extended capture probes (e.g., DNA molecules) can act as templates for an amplification reaction (e.g., a polymerase chain reaction). [000205] Additional variants of spatial analysis methods, including in some embodiments, an imaging step, are described in Section (II)(a) of PCT Publication No. WO2020/176788 and/or 48 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    U.S. Patent Application Publication No.2020/0277663. Analysis of captured analytes (and/or intermediate agents or portions thereof), for example, including sample removal, extension of capture probes, sequencing (e.g., of a cleaved extended capture probe and/or a cDNA molecule complementary to an extended capture probe), sequencing on the array (e.g., using, for example, in situ hybridization or in situ ligation approaches), temporal analysis, and/or proximity capture, is described in Section (II)(g) of PCT Publication No. WO2020/176788 and/or U.S. Patent Application Publication No.2020/0277663. Some quality control measures are described in Section (II)(h) of PCT Publication No. WO2020/176788 and/or U.S. Patent Application Publication No.2020/0277663. [000206] Spatial information can provide information of medical importance. For example, the methods described herein can allow for: identification of one or more biomarkers (e.g., diagnostic, prognostic, and/or for determination of efficacy of a treatment) of a disease or disorder; identification of a candidate drug target for treatment of a disease or disorder; identification (e.g., diagnosis) of a subject as having a disease or disorder; identification of stage and/or prognosis of a disease or disorder in a subject; identification of a subject as having an increased likelihood of developing a disease or disorder; monitoring of progression of a disease or disorder in a subject; determination of efficacy of a treatment of a disease or disorder in a subject; identification of a patient subpopulation for which a treatment is effective for a disease or disorder; modification of a treatment of a subject with a disease or disorder; selection of a subject for participation in a clinical trial; and/or selection of a treatment for a subject with a disease or disorder. Exemplary methods for identifying spatial information of biological and/or medical importance can be found in U.S. Patent Application Publication Nos.2021/0140982, 2021/0198741, and 2021/0199660. [000207] Spatial information can provide information of biological importance. For example, the methods described herein can allow for: identification of transcriptome and/or proteome expression profiles (e.g., in healthy and/or diseased tissue); identification of multiple analyte types in close proximity (e.g., nearest neighbor or proximity based analysis); determination of up- and/or down-regulated genes and/or proteins in diseased tissue; characterization of tumor microenvironments; characterization of tumor immune responses; characterization of cells types and their co-localization in healthy and diseased tissue; and identification of genetic variants 49 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    within tissues (e.g., based on gene and/or protein expression profiles associated with specific disease or disorder biomarkers). [000208] Typically, for spatial array-based methods, a substrate functions as a support for direct or indirect attachment of capture probes to features of the array. A “feature” is an entity that acts as a support or repository for various molecular entities used in spatial analysis. In some embodiments, some or all of the features in an array are functionalized for analyte capture. Exemplary substrates are described in Section (II)(c) of PCT Publication No. WO2020/176788 and/or U.S. Patent Application Publication No.2020/0277663. Exemplary features and geometric attributes of an array can be found in Sections (II)(d)(i), (II)(d)(iii), and (II)(d)(iv) of PCT Publication No. WO2020/176788 and/or U.S. Patent Application Publication No. 2020/0277663. [000209] Generally, analytes and/or intermediate agents (or portions thereof) can be captured when contacting a biological sample with a substrate including capture probes (e.g., a substrate with capture probes embedded, spotted, printed, fabricated on the substrate, or a substrate with features (e.g., beads, wells) comprising capture probes). As used herein, “contact,” “contacted,” and/or “contacting,” a biological sample with a substrate refers to any contact (e.g., direct or indirect) such that capture probes can interact (e.g., bind covalently or non-covalently (e.g., hybridize)) with analytes from the biological sample. Capture can be achieved actively (e.g., using electrophoresis) or passively (e.g., using diffusion). Analyte capture is further described in Section (II)(e) of PCT Publication No. WO2020/176788 and/or U.S. Patent Application Publication No.2020/0277663. [000210] FIG.5 is a schematic diagram showing an exemplary capture probe, as described herein. As shown, the capture probe 502 is optionally coupled to a feature 501 by a cleavage domain 503, such as a disulfide linker. The capture probe can include a functional sequence 504 that are useful for subsequent processing. The functional sequence 504 can include all or a part of sequencer specific flow cell attachment sequence (e.g., a P5 or P7 sequence), all or a part of a sequencing primer sequence, (e.g., a R1 primer binding site, a R2 primer binding site), or combinations thereof. The capture probe can also include a spatial barcode 505. The capture probe can also include a unique molecular identifier (UMI) sequence 506. While FIG.5 shows the spatial barcode 505 as being located upstream (5’) of UMI sequence 506, it is to be 50 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    understood that capture probes wherein UMI sequence 506 is located upstream (5’) of the spatial barcode 505 is also suitable for use in any of the methods described herein. The capture probe can also include a capture domain 507 to facilitate capture of a target analyte. The capture domain can have a sequence complementary to a sequence of a nucleic acid analyte. The capture domain can have a sequence complementary to a connected probe described herein. The capture domain can have a sequence complementary to a capture handle sequence present in an analyte capture agent. The capture domain can have a sequence complementary to a splint oligonucleotide. Such splint oligonucleotide, in addition to having a sequence complementary to a capture domain of a capture probe, can have a sequence complementary to a sequence of a nucleic acid analyte, a portion of a connected probe described herein, a capture handle sequence described herein, and/or a methylated adaptor described herein. [000211] FIG.6 is a schematic illustrating a cleavable capture probe, wherein the cleaved capture probe can enter into a non-permeabilized cell and bind to analytes within the sample. The capture probe 601 contains a cleavage domain 602, a cell penetrating peptide 603, a reporter molecule 604, and a disulfide bond (-S-S-).605 represents all other parts of a capture probe, for example a spatial barcode and a capture domain. [000212] FIG.7 is a schematic diagram of an exemplary multiplexed spatially-barcoded feature. In FIG.7, the feature 701 can be coupled to spatially-barcoded capture probes, wherein the spatially-barcoded probes of a particular feature can possess the same spatial barcode, but have different capture domains designed to associate the spatial barcode of the feature with more than one target analyte. For example, a feature may be coupled to four different types of spatially- barcoded capture probes, each type of spatially-barcoded capture probe possessing the spatial barcode 702. One type of capture probe associated with the feature includes the spatial barcode 702 in combination with a poly(T) capture domain 703, designed to capture mRNA target analytes. A second type of capture probe associated with the feature includes the spatial barcode 702 in combination with a random N-mer capture domain 704 for gDNA analysis. A third type of capture probe associated with the feature includes the spatial barcode 702 in combination with a capture domain complementary to the analyte capture agent of interest 705. A fourth type of capture probe associated with the feature includes the spatial barcode 702 in combination with a capture probe that can specifically bind a nucleic acid molecule 706 that can function in a 51 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    CRISPR assay (e.g., CRISPR/Cas9). While only four different capture probe-barcoded constructs are shown in FIG.7, capture-probe barcoded constructs can be tailored for analyses of any given analyte associated with a nucleic acid and capable of binding with such a construct. For example, the schemes shown in FIG.7 can also be used for concurrent analysis of other analytes disclosed herein, including, but not limited to: (a) mRNA, a lineage tracing construct, cell surface or intracellular proteins and metabolites, and gDNA; (b) mRNA, accessible chromatin (e.g., ATAC-seq, DNase-seq, and/or MNase-seq) cell surface or intracellular proteins and metabolites, and a perturbation agent (e.g., a CRISPR crRNA/sgRNA, TALEN, zinc finger nuclease, and/or antisense oligonucleotide as described herein); (c) mRNA, cell surface or intracellular proteins and/or metabolites, a barcoded labelling agent (e.g., the MHC multimers described herein), and a V(D)J sequence of an immune cell receptor (e.g., T-cell receptor). In some embodiments, a perturbation agent can be a small molecule, an antibody, a drug, an aptamer, a miRNA, a physical environmental (e.g., temperature change), or any other known perturbation agents. [000213] The functional sequences can generally be selected for compatibility with any of a variety of different sequencing systems, e.g., Ion Torrent Proton or PGM, Illumina® sequencing instruments, PacBio, Oxford Nanopore, etc., and the requirements thereof. In some embodiments, functional sequences can be selected for compatibility with non-commercialized sequencing systems. Examples of such sequencing systems and techniques, for which suitable functional sequences can be used, include (but are not limited to) Ion Torrent Proton or PGM sequencing, Illumina® sequencing, PacBio SMRT sequencing, and Oxford Nanopore sequencing. Further, in some embodiments, functional sequences can be selected for compatibility with other sequencing systems, including non-commercialized sequencing systems. [000214] In some embodiments, the spatial barcode 505 and functional sequences 504 is common to all of the probes attached to a given feature. In some embodiments, the UMI sequence 506 of a capture probe attached to a given feature is different from the UMI sequence of a different capture probe attached to the given feature. [000215] FIG.8 depicts an exemplary arrangement of barcoded features within an array. From left to right, FIG.8 shows (L) a slide including six spatially-barcoded arrays, (C) an enlarged schematic of one of the six spatially-barcoded arrays, showing a grid of barcoded features in 52 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    relation to a biological sample, and (R) an enlarged schematic of one section of an array, showing the specific identification of multiple features within the array (labelled as ID578, ID579, ID560, etc.). [000216] In some embodiments, more than one analyte type (e.g., nucleic acids and proteins) from a biological sample can be detected (e.g., simultaneously or sequentially) using any appropriate multiplexing technique, such as those described in Section (IV) of PCT Publication No. WO2020/176788 and/or U.S. Patent Application Publication No.2020/0277663. [000217] In some cases, spatial analysis can be performed by attaching and/or introducing a molecule (e.g., a peptide, a lipid, or a nucleic acid molecule) having a barcode (e.g., a spatial barcode) to a biological sample (e.g., to a cell in a biological sample). In some embodiments, a plurality of molecules (e.g., a plurality of nucleic acid molecules) having a plurality of barcodes (e.g., a plurality of spatial barcodes) are introduced to a biological sample (e.g., to a plurality of cells in a biological sample) for use in spatial analysis. In some embodiments, after attaching and/or introducing a molecule having a barcode to a biological sample, the biological sample can be physically separated (e.g., dissociated) into single cells or cell groups for analysis. Some such methods of spatial analysis are described in Section (III) of PCT Publication No. WO2020/176788 and/or U.S. Patent Application Publication No.2020/0277663. [000218] In some cases, spatial analysis can be performed by detecting multiple oligonucleotides that hybridize to an analyte. In some instances, for example, spatial analysis can be performed using RNA-templated ligation (RTL). Methods of RTL have been described previously. See, e.g., Credle et al., Nucleic Acids Res.2017 Aug 21; 45(14):e128. Typically, RTL includes hybridization of two oligonucleotides to adjacent sequences on an analyte (e.g., an RNA molecule, such as an mRNA molecule). In some instances, the oligonucleotides are DNA molecules. In some instances, one of the oligonucleotides includes at least two ribonucleic acid bases at the 3’ end and/or the other oligonucleotide includes a phosphorylated nucleotide at the 5’ end. In some instances, one of the two oligonucleotides includes a capture binding capture domain (e.g., a poly(A) sequence, a non-homopolymeric sequence). After hybridization to the analyte, a ligase (e.g., a T4 RNA ligase (Rnl2), a PBCV-1 DNA Ligase or Chorella virus DNA Ligase, a single-stranded DNA ligase, or a T4 DNA ligase) ligates the two oligonucleotides together, creating a ligation product. In some instances, the two oligonucleotides hybridize to 53 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    sequences that are not adjacent to one another. For example, hybridization of the two oligonucleotides creates a gap between the hybridized oligonucleotides. In some instances, a polymerase (e.g., a DNA polymerase) can extend one of the oligonucleotides prior to ligation. After ligation, the ligation product is released from the analyte. In some instances, the ligation product is released using an endonuclease (e.g., RNAse H). In some instances, the ligation product is removed using heat. In some instances, the ligation product is removed using KOH. The released ligation product can then be captured by capture probes (e.g., instead of direct capture of an analyte) on an array, optionally amplified, and sequenced, thus determining the location and optionally the abundance of the analyte in the biological sample. [000219] A non-limiting example of templated ligation methods disclosed herein is depicted in FIG.9A. After a biological sample is contacted with a substrate including a plurality of capture probes and contacted with (a) a first probe 901 having a target-hybridization sequence 903 and a primer sequence 902 and (b) a second probe 904 having a target-hybridization sequence 905 and a capture domain (e.g., a poly-A sequence) 906, the first probe 901 and a second probe 904 hybridize 910 to an analyte 907. A ligase 921 ligates 920 the first probe to the second probe thereby generating a ligation product 922. The ligation product is released 930 from the analyte 931 by digesting the analyte using an endoribonuclease 932. The sample is permeabilized 940 and the ligation product 941 is able to hybridize to a capture probe on the substrate. Methods and composition for spatial detection using templated ligation have been described in PCT Publ. No. WO 2021/133849 A1, U.S. Pat. Nos.11,332,790 and 11,505,828, each of which is incorporated by reference in its entirety. [000220] In some embodiments, as shown in FIG.9B, the ligation product 9001 includes a capture probe capture domain 9002, which can bind to a capture probe 9003 (e.g., a capture probe immobilized, directly or indirectly, on a substrate 9004). In some embodiments, methods provided herein include contacting 9005 a biological sample with a substrate 9004, wherein the capture probe 9003 is affixed to the substrate (e.g., immobilized to the substrate, directly or indirectly). In some embodiments, the capture probe capture domain 9002 of the ligated product specifically binds to the capture domain 9006. The capture probe can also include a unique molecular identifier (UMI) 9007, a spatial barcode 9008, a functional sequence 9009, and a cleavage domain 9010. 54 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    [000221] In some embodiments, methods provided herein include permeabilization of the biological sample such that the capture probe can more easily bind to the captured ligated probe (i.e., compared to no permeabilization). In some embodiments, reverse transcription (RT) reagents can be added to permeabilized biological samples. Incubation with the RT reagents can extend the capture probes 9011 to produce spatially-barcoded full-length cDNA 9012 and 9013 from the captured analytes (e.g., polyadenylated mRNA). Second strand reagents (e.g., second strand primers, enzymes) can be added to the biological sample on the slide to initiate second strand synthesis. [000222] In some embodiments, cDNA can be denatured 9014 from the capture probe template and transferred (e.g., to a clean tube) for amplification, and/or library construction. The spatially- barcoded, full-length cDNA can be amplified 9015 via PCR prior to library construction. The cDNA can then be enzymatically fragmented and size-selected in order to optimize the cDNA amplicon size. P59016, i59017, i79018, and P79019, and can be used as sample indexes, and TruSeq™ Read 2 can be added via End Repair, A-tailing, Adaptor Ligation, and PCR. The cDNA fragments can then be sequenced using paired-end sequencing using TruSeq™ Read 1 and TruSeq™ Read 2 as sequencing primer sites. [000223] In some embodiments, detection of one or more analytes (e.g., protein analytes) can be performed using one or more analyte capture agents. As used herein, an “analyte capture agent” refers to an agent that interacts with an analyte (e.g., an analyte in a biological sample) and with a capture probe (e.g., a capture probe attached to a substrate or a feature) to identify the analyte. In some embodiments, the analyte capture agent includes: (i) an analyte binding moiety (e.g., that binds to an analyte), for example, an antibody or antigen-binding fragment thereof; (ii) analyte binding moiety barcode; and (iii) an analyte capture sequence. As used herein, the term “analyte binding moiety barcode” refers to a barcode that is associated with or otherwise identifies the analyte binding moiety. As used herein, the term “analyte capture sequence” refers to a region or moiety configured to hybridize to, bind to, couple to, or otherwise interact with a capture domain of a capture probe. In some cases, an analyte binding moiety barcode (or portion thereof) may be able to be removed (e.g., cleaved) from the analyte capture agent. Additional description of analyte capture agents can be found in Section (II)(b)(ix) of PCT Publication No. 55 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    WO2020/176788 and/or Section (II)(b)(viii) U.S. Patent Application Publication No. 2020/0277663. [000224] FIG.10 is a schematic diagram of an exemplary analyte capture agent 1002 comprised of an analyte-binding moiety 1004 and an analyte-binding moiety barcode domain 1008. The exemplary analyte -binding moiety 1004 is a molecule capable of binding to an analyte 1006 and the analyte capture agent is capable of interacting with a spatially-barcoded capture probe. The analyte -binding moiety can bind to the analyte 1006 with high affinity and/or with high specificity. The analyte capture agent can include an analyte-binding moiety barcode domain 1008, a nucleotide sequence (e.g., an oligonucleotide), which can hybridize to at least a portion or an entirety of a capture domain of a capture probe. The analyte-binding moiety barcode domain 1008 can comprise an analyte binding moiety barcode and a capture handle sequence described herein. The analyte-binding moiety 1004 can include a polypeptide and/or an aptamer. The analyte-binding moiety 1004 can include an antibody or antibody fragment (e.g., an antigen- binding fragment). [000225] FIG.11 is a schematic diagram depicting an exemplary interaction between a feature- immobilized capture probe 1124 and an analyte capture agent 1126. The feature-immobilized capture probe 1124 can include a spatial barcode 1108 as well as functional sequences 1106 and UMI 1110, as described elsewhere herein. The capture probe can be affixed 1104 to a feature (e.g., bead) or array 1102. The capture probe can also include a capture domain 1112 that is capable of binding to an analyte capture agent 1126. The analyte capture agent 1126 can include a functional sequence 1118, analyte binding moiety barcode 1116, and a capture handle sequence 1114 that is capable of binding to the capture domain 1112 of the capture probe 1124. The analyte capture agent can also include a linker 1120 that allows the capture agent barcode domain 1116 to couple to the analyte binding moiety 1122. [000226] During analysis of spatial information, sequence information for a spatial barcode associated with an analyte is obtained, and the sequence information can be used to provide information about the spatial distribution of the analyte in the biological sample. Various methods can be used to obtain the spatial information. In some embodiments, specific capture probes and the analytes they capture are associated with specific locations in an array of features on a substrate. For example, specific spatial barcodes can be associated with specific array 56 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    locations prior to array fabrication, and the sequences of the spatial barcodes can be stored (e.g., in a database) along with specific array location information, so that each spatial barcode uniquely maps to a particular array location. [000227] Alternatively, specific spatial barcodes can be deposited at predetermined locations in an array of features during fabrication such that at each location, only one type of spatial barcode is present so that spatial barcodes are uniquely associated with a single feature of the array. Where necessary, the arrays can be decoded using any of the methods described herein so that spatial barcodes are uniquely associated with array feature locations, and this mapping can be stored as described above. [000228] When sequence information is obtained for capture probes and/or analytes during analysis of spatial information, the locations of the capture probes and/or analytes can be determined by referring to the stored information that uniquely associates each spatial barcode with an array feature location. In this manner, specific capture probes and captured analytes are associated with specific locations in the array of features. Each array feature location represents a position relative to a coordinate reference point (e.g., an array location, a fiducial marker) for the array. Accordingly, each feature location has an “address” or location in the coordinate space of the array. [000229] Some exemplary spatial analysis workflows are described in the Exemplary Embodiments section of PCT Publication No. WO2020/176788 and/or U.S. Patent Application Publication No.2020/0277663. See, for example, the Exemplary embodiment starting with “In some non-limiting examples of the workflows described herein, the sample can be immersed…” of PCT Publication No. WO2020/176788 and/or U.S. Patent Application Publication No. 2020/0277663. See also, e.g., the Visium Spatial Gene Expression Reagent Kits User Guide (e.g., Rev F, dated January 2022); and/or the Visium Spatial Gene Expression Reagent Kits - Tissue Optimization User Guide (e.g., Rev E, dated February 2022). [000230] In some embodiments, spatial analysis can be performed using dedicated hardware and/or software, such as any of the systems described in Sections (II)(e)(ii) and/or (V) of PCT Publication No. WO2020/176788 and/or U.S. Patent Application Publication No.2020/0277663, or any of one or more of the devices or methods described in Sections Control Slide for Imaging, Methods of Using Control Slides and Substrates for, Systems of Using Control Slides and 57 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    Substrates for Imaging, and/or Sample and Array Alignment Devices and Methods, Informational labels of PCT Publication No. WO2020/123320. [000231] Suitable systems for performing spatial analysis can include components such as a chamber (e.g., a flow cell or sealable, fluid-tight chamber) for containing a biological sample. The biological sample can be mounted for example, in a biological sample holder. One or more fluid chambers can be connected to the chamber and/or the sample holder via fluid conduits, and fluids can be delivered into the chamber and/or sample holder via fluidic pumps, vacuum sources, or other devices coupled to the fluid conduits that create a pressure gradient to drive fluid flow. One or more valves can also be connected to fluid conduits to regulate the flow of reagents from reservoirs to the chamber and/or sample holder. [000232] The systems can optionally include a control unit that includes one or more electronic processors, an input interface, an output interface (such as a display), and a storage unit (e.g., a solid state storage medium such as, but not limited to, a magnetic, optical, or other solid state, persistent, writeable and/or re-writeable storage medium). The control unit can optionally be connected to one or more remote devices via a network. The control unit (and components thereof) can generally perform any of the steps and functions described herein. Where the system is connected to a remote device, the remote device (or devices) can perform any of the steps or features described herein. The systems can optionally include one or more detectors (e.g., CCD, CMOS) used to capture images. The systems can also optionally include one or more light sources (e.g., LED-based, diode-based, lasers) for illuminating a sample, a substrate with features, analytes from a biological sample captured on a substrate, and various control and calibration media. [000233] The systems can optionally include software instructions encoded and/or implemented in one or more of tangible storage media and hardware components such as application specific integrated circuits. The software instructions, when executed by a control unit (and in particular, an electronic processor) or an integrated circuit, can cause the control unit, integrated circuit, or other component executing the software instructions to perform any of the method steps or functions described herein. [000234] In some cases, the systems described herein can detect (e.g., register an image) the biological sample on the array. Exemplary methods to detect the biological sample on an array 58 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    are described in PCT Publication No. WO2021/102003 and/or U.S. Patent Application Publication No.2021/0150707, each of which is incorporated herein by reference in their entireties. [000235] Prior to transferring analytes from the biological sample to the array of features on the substrate, the biological sample can be aligned with the array. Alignment of a biological sample and an array of features including capture probes can facilitate spatial analysis, which can be used to detect differences in analyte presence and/or level within different positions in the biological sample, for example, to generate a three-dimensional map of the analyte presence and/or level. Exemplary methods to generate a two- and/or three-dimensional map of the analyte presence and/or level are described in PCT Publication No. WO2020/053655 and spatial analysis methods are generally described in PCT Publication No. WO2021/102039 and/or U.S. Patent Application Publication No.2021/0155982, each of which is incorporated herein by reference in their entireties. [000236] In some cases, a map of analyte presence and/or level can be aligned to an image of a biological sample using one or more fiducial markers, e.g., objects placed in the field of view of an imaging system which appear in the image produced, as described in the Substrate Attributes Section, Control Slide for Imaging Section of PCT Publication Nos. WO2020/123320, WO 2021/102005, and/or U.S. Patent Application Publication No.2021/0158522, each of which is incorporated herein by reference in their entireties. Fiducial markers can be used as a point of reference or measurement scale for alignment (e.g., to align a sample and an array, to align two substrates, to determine a location of a sample or array on a substrate relative to a fiducial marker) and/or for quantitative measurements of sizes and/or distances. III. ENGINEERED REVERSE TRANSCRIPTASES [000237] Reverse transcriptases or reverse transcription (RT) enzymes are RNA-dependent DNA polymerases, typically used to create a copy of an RNA sequence thereby generating a cDNA molecule. Reverse transcription is initiated by hybridization of a priming sequence to an RNA molecule which is extended by a reverse transcription enzyme in a template directed fashion. A reverse transcription enzyme adds a plurality of non-template nucleotides to a nucleotide strand, thereby producing complementary deoxyribonucleic acid (cDNA) molecules. The resultant cDNA can then be dehybridized from the template RNA molecule in any number of ways as 59 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    known in the art. Engineered and/or recombinant are used interchangeably with respect to reverse transcriptase (RT) variant and/or fusion RT. [000238] One aspect of the present disclosure provides an engineered reverse transcriptase (RT) polypeptide comprising an RT polypeptide sequence; a DNA binding domain, and a linker connecting the RT polypeptide sequence and the DNA binding domain. The DNA binding domain can be from a molecule capable of binding a minor groove of a nucleic acid (e.g., RNA or DNA). [000239] Another aspect of the present disclosure provides a recombinant reverse transcriptase (RT) protein comprising a RT polypeptide fused to a DNA binding domain. In some embodiments, the RT polypeptide and the DNA binding domain are separated by an amino acid linker, and the DNA binding domain is fused to the C-terminus of the RT polypeptide. In some embodiments, the RT polypeptide and the DNA binding domain are separated by an amino acid linker, and the DNA binding domain is fused to the N-terminus of the RT polypeptide. A. DNA Binding domains [000240] A DNA binding domain is a protein, or a defined region of a protein, that binds to a nucleic acid in a sequence-independent matter. For example, binding of the protein to DNA does not exhibit any preference for a particular sequence. The DNA binding domain may be single or double stranded. The nucleic acid binding domain can comprise a single stranded DNA binding protein; a double stranded DNA binding protein; a single stranded RNA binding protein; a double stranded RNA binding protein; a continuous RNA-DNA hybrid binding protein; or a discontinuous RNA-DNA hybrid binding protein. [000241] The nucleic acid binding domain can help stabilize the interaction between the RNA template and the DNA primer during reverse transcription. For example, the nucleic acid binding domain can enhance the efficiency and/or processivity of the engineered RT polypeptide during reverse transcription. Suitable DNA binding domains of the present disclosure can be identical to or substantially identical to a known DNA binding protein over a comparison window of about 25 amino acids, about 50 to about 100 amino acids, any value in-between these two parameters of 25 and 100 amino acids (e.g., about 55 to about 75 amino acids), or over the length of the entire protein. The sequence can be compared and aligned for maximum correspondence over a 60 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    comparison window, or designated region as measured using one of the described comparison algorithms or by manual alignment and visual inspection. For purposes of this disclosure, percent amino acid identity is determined by the default parameters of BLAST and or CLUSTAL W. [000242] DNA binding domain (DBD) proteins or polypeptides are capable of binding DNA. DNA binding domains may include, but are not limited to, one or more DNA binding domains from an archaeal DNA binding protein, single-stranded DNA binding domains and/or 7 kDa DNA binding domains. The DNA binding domain can be a DNA binding domain of any one of Saccharomyces cerevisiae datin (DAT1), high mobility group AT hook 1 (HMGA1), lysine- specific methyltransferase 2a ( KMT2A), Myocyte Enhancer Factor 2C (MEF2C), Heterogeneous Nuclear Ribonucleoprotein D (HNRNPD), Structural Maintenance of Chromosomes 1A (SMC1), Structural Maintenance Of Chromosomes 2 (SMC2), Caenorhabditis elegans tbp-1, Drosophila melanogaster D1 protein, Salmonella typhimurium Hin recombinase, S. typhimurium Gin recombinase, S. typhimurium Pin recombinase, or S. typhimurium Cin recombinase, or a combination thereof. The sequence specific DNA binding protein of any one of these molecules can be altered to produce an engineered RT polypeptide or a recombinant RT protein disclosed herein. Any proteins having substantially the same function as DAT1 and comprising any peptide motifs with sequence specific DNA binding function can be used to engineer the recombinant RT protein or engineered RT polypeptide described herein. [000243] Specifically, the DNA binding domain can be from a S. cerevisiae DAT1. DAT1 is a yeast protein (e.g., Saccharomyces cerevisiae) that specifically recognizes the minor groove of non-alternating oligo(A)-oligo(T) tracts (e.g., >10 bp oligo(A)-oligo(T) tract). See e.g., Reardon et al., PNAS 90, 11327 (1993); Reardon et al. Nucleic Acids Research, 23, 4900 (1995). In some embodiments of the present disclosure, the DNA binding domain can comprise SEQ ID NO: 2. In some embodiments, the DNA binding domain is encoded by SEQ ID NO: 25. In some embodiments, the amino acid sequence of the DNA binding domain comprises a DNA binding domain consensus motif set forth in SEQ ID NO: 13, 14, 16, or 22. [000244] The sequence specific recognition may be determined by to three repeated pentads of G-R-K-P-G (SEQ ID NO: 11). Accordingly, in some embodiments of the present disclosure, the DNA binding domain comprises the amino acid sequence of SEQ ID NO: 11 (GRKPG). Alternatively, the DNA binding domain can comprise at least 2 domains, at least 3 domains, at 61 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    least 4 domains, at least five domains, at least six domains, at least seven domain, at least eight domains, at least nine domain, or at least ten domain comprising SEQ ID NO: 11. [000245] In some embodiments, the DNA binding domain can specifically recognize adenine- thymine-rich region on a nucleic acid molecule. The DNA binding domain can specifically recognize oligo(dA) or oligo(dT) tracts on a nucleic acid molecule. The DNA binding domain can comprise or consist of at least one AT-rich interaction domain. For example, the DNA binding domain can comprise; consist of , or consist essentially of at least 2, at least 3, at least 4, at least 5, at least 6, at least 7, at least 8, at least 9, or at least 10 AT-rich interaction domains. The AT-rich interaction domain can comprise a core sequence. This core sequence can be a two- base core sequence, a three-base core sequence, a four-base core sequence, or a five-base core sequence. In some embodiments, the at least one of the bases of the core sequence can comprise an arginine. [000246] In some embodiments, the at least one of the bases of the core sequence can comprise a glycine and an arginine. In some embodiments, the at least one of the bases of the core sequence can comprise a proline and an arginine. In some embodiments, the at least one of the bases of the core sequence can comprise a lysine and an arginine. or any combination thereof. In some embodiments, the at least one of the bases of the core sequence comprises an arginine; a glycine and an arginine; a proline and an arginine; a lysine and an arginine; or any combination thereof. [000247] The AT-rich interaction domain contemplated by the present disclosure comprises a GRKPG (Gly-Arg-Lys-Pro-Gly) repeat, a RKRGRPKK repeat, a KKRGRPKK repeat, a RKRGR repeat, a GR*R/PPK repeat, a GR*RPK repeat, a GR*PPK repeat, a KRPR* repeat, or a K/RKRGRPKK repeat. In some embodiments, the AT-rich interaction domain can be a GRKPG (Gly-Arg-Lys-Pro-Gly) repeat or SEQ ID NO: 11. In some embodiments, the AT-rich interaction domain can be a RKRGRPKK repeat or SEQ ID NO: 16 or 17. In some embodiments, the AT-rich interaction domain can be KKRGRPKK repeat or SEQ ID NO: 18. In some embodiments, the AT-rich interaction domain can be a RKRGR repeat or SEQ ID NO; 19. In some embodiments, the AT-rich interaction domain can be a GR*R/PPK repeat or SEQ ID NO: 22. In some embodiments, the AT-rich interaction domain can be a GR*RPK repeat, or SEQ ID NO: 23. In some embodiments, the AT-rich interaction domain can be a GR*PPK repeat, or SEQ ID NO: 24. In some embodiments, the AT-rich interaction domain can be a 62 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    KRPR* repeat, or SEQ ID NO: 21. In some embodiments, the AT-rich interaction domain can be a K/RKRGRPKK repeat, or SEQ ID NO: 16. [000248] In some embodiments, the AT-rich interaction domain can be a mammalian high mobility group I protein (HMG-I, or a-protein) AT-rich interaction domain or SEQ ID NO: 15. In some embodiments, the AT-rich interaction domain can comprise a Drosophila melanogaster D1 protein AT-rich interaction domain-consensus domain or SEQ ID NO: 16. In some embodiments, the AT-rich interaction domain can be a Drosophila melanogaster D1 protein AT- rich interaction domain or SEQ ID NO: 17 or SEQ ID NO: 18. In some embodiments, the AT- rich interaction domain can comprise a phage 434 repressor AT-rich interaction domain or SEQ ID NO: 21. In some embodiments, the AT-rich interaction domain can comprise a Hin recombinase of Salmonella typhimurium AT-rich interaction domain or SEQ ID NO: 22, SEQ ID NO: 23, or SEQ ID NO: 24. In some embodiments, the AT-rich interaction domain comprises a core sequence comprising an amino acid selected from the group consisting of SEQ ID NO: 11- 24. [000249] In some embodiments, the DNA binding domain is a S. cerevisiae datin (DAT1) DNA binding domain or fragment thereof. Several fragments were tested and shown to maintain DNA binding activity. For example, the N-terminal 90 amino acids (D90) and/or the N-terminal 36 amino acids (D36) bind in a sequence specific manner to oligo(A)-oligo(T) tract. DAT1(D-90) can specifically bind to A-T tracts with Kd of about 3 x 10-10 M (or 3 x 10-9 M); and DAT1(D-36) protein can bind to A-T tracts with Kd of 4 x 10-10 M. DAT1(D-90) can also be more resistant to degradation by bacterial proteases than longer and shorter DAT1 derivatives. The DNA binding activity of DAT1(D-90) was also resistant to heat (boiling in water bath for 10 min) and chemical treatment (6 M guanidine HCl). DAT1(D-90) was also shown to be highly soluble in physiologic salt and pH conditions. [000250] Accordingly, in some embodiments of the engineered RT polypeptide disclosed herein, the DNA binding domain can comprise a full-length DAT1 sequence. In some embodiments, the DNA binding domain comprises SEQ ID NO: 2. The DNA binding domain can also comprise an N-terminal truncated variant of DAT1 (D90) comprising the first 90 amino acids of the full length DAT1. In that embodiment, the DNA binding domain can comprise SEQ ID NO: 3. 63 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    [000251] The DNA binding domain can also comprise a truncated variant of DAT1 (D60) comprising the first 60 amino acids of the full length DAT1. In that embodiment, the DNA binding domain can comprise SEQ ID NO: 5. The DNA binding domain can also comprise a truncated variant of DAT1 (D48) comprising the first 48 amino acids of the full length DAT1. In that embodiment, the DNA binding domain can comprise SEQ ID NO: 6. The DNA binding domain can also comprise a truncated variant of DAT1(D36) comprising the first 36 amino acids of the full length DAT1. In that embodiment, the DNA binding domain can comprise SEQ ID NO: 8. The DNA binding domain can also comprise a truncated variant of DAT1 (D35) comprising the first 35 amino acids of full length DAT1. In that embodiment, the DNA binding domain can comprise SEQ ID NO: 9. [000252] In some embodiments, the DNA binding domain can comprise an amino acid sequence having at least about 90%, at least about 91%, at least about 92%, at least about 93%, at least about 94%, at least about 95%, at least about 96%, at least about 97%, at least about 98%, or at least about 99% sequence identity to any one of SEQ ID NO: 2, 3, 5, 6, 8, or 9. In some embodiments, the DNA binding domain can also comprise an amino acid sequence having 90%, 91%, 92%, 93%, 94%, 95%, 96%, 97%, 98%, or 99% sequence identity to any one of SEQ ID NO: 2, 3, 5, 6, 8, or 9. In some embodiments of the engineered RT polypeptide or recombinant RT protein described herein, the DNA binding domain can comprise the amino acid sequence of SEQ ID NO: 3, 8, or 9. [000253] In some embodiments of the engineered RT polypeptide or recombinant RT protein described herein, the DNA binding domain can comprise a mutation in any of one of SEQ ID NO: 2, 3, 5, 6, 8, 9, or 11. The mutation can be selected from a substitution, an insertion, a deletion, or any combination thereof. In some embodiments, the mutation can further enhance the sensitivity and/or performance of the engineered RT polypeptide as described herein. [000254] One aspect of the present disclosure provides an engineered reverse transcriptase (RT) polypeptide comprising an RT polypeptide sequence; at least two DNA binding domains, and a linker connecting the RT polypeptide sequence and the at least two DNA binding domains. In that embodiment, each DNA binding domain can be from a molecule capable of binding a minor groove of a nucleic acid. When two DNA binding domains are present, at least one DNA binding domain can be located at the N-terminus of the engineered RT and at least one DNA binding 64 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    domain can be located at the C-terminus of the engineered RT. Alternatively, the at least two DNA binding domains can both be located at the C-terminus or N-terminus of the engineered RT. In some embodiments, the at least two DNA binding domains can be derived from the same molecule. In some embodiments, the at least two DNA binding domains can be derived from different molecules. In some embodiments, the at least two DNA binding domains can be derived from the same organism. In some embodiments, the at least two DNA binding domains can be derived from different organisms. B. Linkers [000255] The engineered reverse transcriptase (RT) polypeptide or the recombinant RT protein described herein comprises a linker. The linker connects the RT polypeptide sequence and the DNA binding domain. Any functional linker known in the art can be used. Any suitable linker, including without limitation any variation of G(n)S(m)G(p) linker can be inserted between the RT polypeptide and the DNA binding protein. In this embodiment, n=0, 1, 2, 3, 4, 5,6 , 7, 8, 9, 1011, 12, 13, 14, 15, 16, 17, 18, 19 or 20, m=0,1, 2, 3, 4, 5,6 , 7, 8, 9, 1011, 12, 13, 14, 15, 16, 17, 18, 19 or 20, p=0, 1, 2, 3, 4, 5,6 , 7, 8, 9, 1011, 12, 13, 14, 15, 16, 17, 18, 19 or 20, and n, m, and p are selected independently. In some embodiments, the linker can be a glycine-serine linker selected from the group consisting of (GS)n, (GSGGS)n, (SGGSG)n, (GGGS)n, (GGSG)n, (GGSGG)n, (GSGSG)n, (GSGGG)n, GGGSG)n, and (GSSSG)n, where n represents an integer of at least 1. In some embodiments, the linker can be GGGS. In some embodiments, the linker can be SGGSG. The linker can comprise GGGGS or SEQ ID NO: 26. The linker can comprise GSGGSG or SEQ ID NO: 199. In some embodiments, the linker can be encoded by SEQ ID NO: 200 or GGTTCAGGGGGTTCCGGT. [000256] The DNA binding domain described herein can be located at the N-terminus of the RT polypeptide sequence. The DNA binding domain described herein can be located at the C- terminus of the RT polypeptide sequence. C. Tag Proteins [000257] One aspect of the present disclosure provides an engineered reverse transcriptase (RT) polypeptide comprising an RT polypeptide sequence; a DNA binding domain from a molecule capable of binding a minor groove of a nucleic acid; a linker connecting the RT 65 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    polypeptide sequence and the DNA binding domain; and a tag protein. [000258] Another aspect of the present disclosure provides a recombinant reverse transcriptase (RT) protein comprising a RT polypeptide, fused to a DNA binding domain and a tag protein. In that embodiment, the RT polypeptide and the DNA binding domain are separated by an amino acid linker, and the DNA binding domain is fused to the C-terminus of the RT polypeptide. In that embodiment, the RT polypeptide and the DNA binding domain are separated by an amino acid linker, and the DNA binding domain is fused to the N-terminus of the RT polypeptide [000259] The tag protein can be selected from the group consisting of an affinity tag, a fluorescent tag, or an expression and/or solubility enhancement tag. In some embodiments, the tag protein is selected from hexahistidine tag (his-tag), Fasciola hepatica 8-kDa antigen tag (Fh8), Glutathione-S-transferase (GST) tag, maltose-binding protein tag (MBP), Flag tag peptide (FLAG tag), streptavidin binding peptide tag (Strep-II), calmodulin-binding protein tag (CBP), mutated dehalogenase tag (HaloTag), staphylococcal Protein A (Protein A), intein mediated purification with the chitin-binding domain (IMPACT (CBD)), cellulose-binding module (CBM), dockerin domain of Clostridium josui tag (Dock), fungal avidin-like protein (Tamavidin), small ubiquitin-like modifier tag ( SUMO), a strep tag, Thioredoxin (Trx) tag, a VariFlex C-Terminal solubility enhancement tag, a short peptide C-terminal tag, Solubility- enhancer peptide sequences (SET) tag, IgG domain B1 of Protein G (GB1) tag, IgG repeat domain ZZ of Protein A (ZZ) tag, Mutated dehalogenase tag (HaloTag), Solubility eNhancing Ubiquitous Tag (SNUT tag), Seventeen kilodalton protein (Skp tag), Phage T7 protein kinase (T7PK) tag, E. coli secreted protein A (EspA) tag, Monomeric bacteriophage T70.3 protein (Orc protein) (Mocr) tag, E. coli trypsin inhibitor (Ecotin) tag, Calcium-binding protein (CaBP) tag, Stress-responsive arsenate reductase (ArsC) tag, N-terminal fragment of translation initiation factor IF2 (IF2-domain I) tag, N-terminal fragment of translation initiation factor IF2 (Expressivity) tag, Stress-responsive proteins tag (e.g., RpoA, tag, SlyD Tsf tag, RpoS tag, PotD tag, or Crr tag), and E. coli acidic proteins tag (e.g., msyB tag, yigD tag, and rpoD tag). Additional affinity tags and solubility enhancer tags are known to those skill in the art. See Costa et al., Front. Microbiol., 63(5): (2014); Esposito and Chatterjee Curr. Opin. Biotechnol., 17: 353–358 (2006); Malhotra, A. “Tagging for protein expression,” in Guide to Protein Purification, 66 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    2nd Edn, eds. R. R. Burgess and M. P. Deutscher (San Diego, CA: Elsevier), 463:239–258 (2009). [000260] In some embodiments, the tag is selected from hexahistidine tag (his-tag), small ubiquitin-like modifier tag (SUMO), a VariFlex C-Terminal solubility enhancement tag, a short peptide C-terminal tag, Thioredoxin (Trx) tag, Solubility-enhancer peptide sequences (SET) tag, IgG domain B1 of Protein G (GB1) tag, IgG repeat domain ZZ of Protein A (ZZ) tag, Solubility enhancing Ubiquitous Tag (SNUT tag), Seventeen kilodalton protein (Skp tag), Phage T7 protein kinase (T7PK) tag, E. coli secreted protein A (EspA) tag, Monomeric bacteriophage T70.3 protein (Orc protein) (Mocr) tag, E. coli trypsin inhibitor (Ecotin) tag, Calcium-binding protein (CaBP) tag, Stress-responsive arsenate reductase (ArsC) tag, N-terminal fragment of translation initiation factor IF2 (IF2-domain I) tag, N-terminal fragment of translation initiation factor IF2 (Expressivity) tag, Fasciola hepatica 8-kDa antigen tag (Fh8), Glutathione-S-transferase (GST) tag, maltose-binding protein tag (MBP), Flag tag peptide (FLAG), streptavidin binding peptide tag (Strep-II; strep), calmodulin-binding protein tag (CBP), mutated dehalogenase tag (HaloTag), staphylococcal Protein A (Protein A), intein mediated purification with the chitin-binding domain (IMPACT (CBD)), cellulose-binding module (CBM), dockerin domain of Clostridium josui tag (Dock), or fungal avidin-like protein (Tamavidin). [000261] Tags used in the practice of any invention disclosed herein may serve any number of purposes and a number of tags may be added to impart one or more different functions to the engineered reverse transcriptase, and/or derivatives thereof, of the disclosure. For example, tags may (1) contribute to protein-protein interactions both internally within a protein and with other protein molecules, (2) make the protein amenable to particular purification methods, (3) enable one to identify whether the protein is present in a composition; or (4) give the protein other functional characteristics. [000262] In one embodiment, the tag is an affinity tag selected from a histidine tag such as a hexahistidine tag (his-tag or 6 His-tag), Fasciola hepatica 8-kDa antigen tag (Fh8), Glutathione- S-transferase (GST) tag, maltose-binding protein tag (MBP), Flag tag peptide (FLAG), streptavidin binding peptide tag (Strep-II), calmodulin-binding protein tag (CBP), mutated dehalogenase tag (HaloTag), staphylococcal Protein A (Protein A), intein mediated purification with the chitin-binding domain (IMPACT (CBD)), cellulose-binding module (CBM), dockerin 67 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    domain of Clostridium josui tag (Dock), or fungal avidin-like protein (Tamavidin). In one embodiment, the tag is a hexahistidine tag. [000263] In some embodiments, the tag is selected from a small ubiquitin-like modifier tag (SUMO), a VariFlex C-Terminal solubility enhancement tag, a short peptide C-terminal tag, Thioredoxin (Trx) tag, Solubility-enhancer peptide sequences (SET) tag, IgG domain B1 of Protein G (GB1) tag, IgG repeat domain ZZ of Protein A (ZZ) tag, Solubility enhancing Ubiquitous Tag (SNUT tag), Seventeen kilodalton protein (Skp tag), Phage T7 protein kinase (T7PK) tag, E. coli secreted protein A (EspA) tag, Monomeric bacteriophage T70.3 protein (Orc protein) (Mocr) tag, E. coli trypsin inhibitor (Ecotin) tag, Calcium-binding protein (CaBP) tag, Stress-responsive arsenate reductase (ArsC) tag, N-terminal fragment of translation initiation factor IF2 (IF2-domain I) tag, N-terminal fragment of translation initiation factor IF2 (Expressivity) tag, Fasciola hepatica 8-kDa antigen tag (Fh8), Glutathione-S-transferase (GST) tag, maltose-binding protein tag (MBP), Flag tag peptide (FLAG), streptavidin binding peptide tag (Strep-II; strep), calmodulin-binding protein tag (CBP), mutated dehalogenase tag (HaloTag), staphylococcal Protein A (Protein A), intein mediated purification with the chitin-binding domain (IMPACT (CBD)), cellulose-binding module (CBM), dockerin domain of Clostridium josui tag (Dock), fungal avidin-like protein (Tamavidin). [000264] In some embodiments, the solubility enhancer tag is selected from the group consisting of a SUMO tag, a GST tag, a Trx tag, a VariFlex C-Terminal solubility enhancement tag, a short peptide C-terminal tag, an Fh8 tag, MBP tag, SET tag, GB1 tag, ZZ tag, HaloTag, SNUT tag, Skp tag, T7PK tag, EspA tag, Mocr tag, Ecotin tag, CaBO tag, ArsC tag, IF2-domain I tag, Expressivity tag, RpoA, tag, SlyD, tag, Tsf tag, RpoS tag, PotD tag, Crr tag, msyB tag, yigD tag, and rpoD tag. [000265] In some embodiments, the tag is an affinity tag. In one embodiment, the tag is an affinity tag and comprises a histidine purification tag. In one embodiment, the tag is a hexahistidine tag (his tag). In one embodiment, the tag comprises an amino acid sequence of the sequence HHHHHH (SEQ ID NO: 62). In one embodiment, the tag is a solubility enhancer tag. In one embodiment, the solubility enhancer tag is a short peptide C-terminal tag. In one embodiment, the solubility enhancer tag comprises an amino acid sequence of SEEDEEKEEDG 68 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    (SEQ ID NO: 193) or an amino acid sequence having at least 90% sequence identity to the amino acid sequence of SEQ ID NO: 193. [000266] In some embodiments, the tag further comprises an endoprotein cleavage site selected from ENLYFQ/G (SEQ ID NO: 194), DDDDK/ (SEQ ID NO: 195), IEGR/ (SEQ ID NO: 196), LVPR/GS (SEQ ID NO: 197), or LEVLFQ/GP (SEQ ID NO: 198). [000267] In some embodiments, the engineered nucleic acid processing enzyme or a derivative thereof further comprises a protease cleavage sequence. In some embodiments, the cleavage of the protease cleavage sequence by a protease results in cleavage of the affinity tag from the engineered reverse transcriptase polypeptide or recombinant RT protein or a derivative thereof. In some instances, the protease cleavage sequence/site is recognized by a protease including, but not limited to, alanine carboxypeptidase, Armillaria mellea astacin, bacterial leucyl aminopeptidase, cancer procoagulant, cathepsin B, clostripain, cytosol alanyl aminopeptidase, elastase, endoproteinase Arg-C, enterokinase (EnTK), gastricsin, gelatinase, Gly-X carboxypeptidase, glycyl endopeptidase, human rhinovirus 3C protease, hypodermin C, Iga- specific serine endopeptidase, leucyl aminopeptidase, leucyl endopeptidase, lysC, lysosomal pro- X carboxypeptidase, lysyl aminopeptidase, methionyl aminopeptidase, myxobacter, nardilysin, pancreatic endopeptidase E, picornain 2A, picornain 3C, proendopeptidase, prolyl aminopeptidase, proprotein convertase I, proprotein convertase II, russellysin, saccharopepsin, semenogelase, T-plasminogen activator, thrombin (Thr), tissue kallikrein, tobacco etch virus (TEV), togavirin, tryptophanyl aminopeptidase, U-plasminogen activator, V8, venombin A, venombin AB, factor Xa (Xa), and Xaa-pro aminopeptidase. In some embodiments, the protease cleavage sequence is a thrombin cleavage sequence. [000268] In some embodiments, the tag is cleaved or removed from the engineered nucleic acid processing enzyme or derivatives thereof via the cleavage site. In one embodiment, the tag is cleaved or removed using an endoprotein selected from the group consisting of tobacco etch virus protease (Tev), enterokinase (EntK), factor Xa (Xa), thrombin (Thr), genetically engineered derivative of human rhinovirus 3C protease (PreScission), Catalytic core of Ulp1 (SUMO protease). In one embodiment, the tag is cleaved at ENLYFQ/G (SEQ ID NO: 194) using tobacco etch virus protease (Tev). In another embodiment, the tag is cleaved at DDDDK/ (SEQ ID NO: 195) using Enterokinase (EntK). In another embodiment, the tag is cleaved at 69 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    IEGR/ (SEQ ID NO: 196) using Factor Xa (Xa). In another embodiment, the tag is cleaved at LVPR/GS (SEQ ID NO: 197) using thrombin (Thr). In another embodiment, the tag is cleaved at LEVLFQ/GP (SEQ ID NO: 198) using a genetically engineered derivative of human rhinovirus 3C protease. In another embodiment, the tag is cleaved with Catalytic core of Ulp1 (SUMO protease). Catalytic core of Ulp1 recognizes SUMO tertiary structure and cleaves at the C- terminal end of the conserved Gly–Gly sequence in SUMO. [000269] In some embodiments, the engineered RT polypeptide, the recombinant RT protein, or derivatives thereof comprises an affinity tag at the N-terminus or at the C-terminus of the amino acid sequence. In some embodiments, the affinity tag include, but is not limited to, albumin binding protein (ABP), AU1 epitope, AU5 epitope, T7-tag, V5-tag, B-tag, Chloramphenicol Acetyl Transferase (CAT), Dihydrofolate reductase (DHFR), AviTag, Calmodulin-tag, polyglutamate tag, E-tag, FLAG-tag, HA-tag, Myc-tag, NE-tag, S-tag, SBP-tag, Doftag 1, Softag 3, Spot-tag, tetracysteine (TC) tag, Ty tag, VSV-tag, Xpress tag, biotin carboxyl carrier protein (BCCP), green fluorescent protein tag, HaloTag, Nus-tag, thioredoxin-tag, Fc-tag, cellulose binding domain, chitin binding protein (CBP), choline-binding domain, galactose binding domain, maltose binding protein (MBP), Horseradish Peroxidase (HRP), Strep-tag, HSV epitope, Ketosteroid isomerase (KSI), KT3 epitope, LacZ, Luciferase, PDZ domain, PDZ ligand, Polyarginine (Arg-tag), Polyaspartate (Asp-tag), Polycysteine (Cys-tag), Polyphenylalanine (Phe-tag), Profinity eXact, Protein C, S1-tag, S1-tag, Staphylococcal protein A (Protein A), Staphylococcal protein G (Protein G), Small Ubiquitin-like Modifier (SUMO), Tandem Affinity Purification (TAP), TrpE, Ubiquitin, Universal, glutathione-S-transferase (GST), and poly(His) tag. In some instances, the affinity tag is at least 5 histidine amino acids. [000270] In some embodiments, engineered reverse transcription enzymes, engineered reverse transcriptases, engineered reverse transcriptase polypeptides, or recombinant RT proteins described herein may comprise an affinity tag at the N-terminus or at a C-terminus of the amino acid sequence. In some instances, the affinity tag may include, but is not limited to, albumin binding protein (ABP), AU1 epitope, AU5 epitope, T7-tag, V5-tag, B-tag, Chloramphenicol Acetyl Transferase (CAT), Dihydrofolate reductase (DHFR), AviTag, Calmodulin-tag, polyglutamate tag, E-tag, FLAG-tag, HA-tag, Myc-tag, NE-tag, S-tag, SBP-tag, Doftag 1, Softag 3, Spot-tag, tetracysteine (TC) tag, Ty tag, VSV-tag, Xpress tag, biotin carboxyl carrier protein 70 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    (BCCP), green fluorescent protein tag, HaloTag, Nus-tag, thioredoxin-tag, Fc-tag, cellulose binding domain, chitin binding protein (CBP), choline-binding domain, galactose binding domain, maltose binding protein (MBP), Horseradish Peroxidase (HRP), Strep-tag, HSV epitope, Ketosteroid isomerase (KSI), KT3 epitope, LacZ, Luciferase, PDZ domain, PDZ ligand, Polyarginine (Arg-tag), Polyaspartate (Asp-tag), Polycysteine (Cys-tag), Polyphenylalanine (Phe-tag), Profinity eXact, Protein C, S1-tag, S1-tag, Staphylococcal protein A (Protein A), Staphylococcal protein G (Protein G), Small Ubiquitin-like Modifier (SUMO), Tandem Affinity Purification (TAP), TrpE, Ubiquitin, Universal, glutathione-S-transferase (GST), and poly(His) tag. In some instances, said affinity tag is at least 6 histidine amino acids (SEQ ID NO: 26). [000271] In some embodiments, an engineered reverse transcriptase polypeptide and/or a recombinant reverse transcriptase protein described herein can comprise a protease cleavage sequence. In that embodiment, cleavage by a protease results in cleavage of the affinity tag from the engineered reverse transcription enzyme. In some instances, the protease cleavage sequence is recognized by a protease including, but not limited to, alanine carboxypeptidase, Armillaria mellea astacin, bacterial leucyl aminopeptidase, cancer procoagulant, cathepsin B, clostripain, cytosol alanyl aminopeptidase, elastase, endoproteinase Arg-C, enterokinase, gastricsin, gelatinase, Gly-X carboxypeptidase, glycyl endopeptidase, human rhinovirus 3C protease, hypodermin C, Iga-specific serine endopeptidase, leucyl aminopeptidase, leucyl endopeptidase, lysC, lysosomal pro-X carboxypeptidase, lysyl aminopeptidase, methionyl aminopeptidase, myxobacter, nardilysin, pancreatic endopeptidase E, picornain 2A, picornain 3C, proendopeptidase, prolyl aminopeptidase, proprotein convertase I, proprotein convertase II, russellysin, saccharopepsin, semenogelase, T-plasminogen activator, thrombin, tissue kallikrein, tobacco etch virus (TEV), togavirin, tryptophanyl aminopeptidase, U-plasminogen activator, V8, venombin A, venombin AB, and Xaa-pro aminopeptidase. In some instances, the protease cleavage sequence is a thrombin cleavage sequence. D. Reverse Transcriptase Polypeptides [000272] Reverse transcriptases or reverse transcription enzymes are known in the art to perform a reverse transcription reaction. As used herein, “Reverse transcriptase” and “reverse transcription enzyme” are synonymous. Reverse transcription is initiated by hybridization of a priming sequence to an RNA molecule which is extended by an engineered reverse transcription 71 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    enzyme in a template directed fashion. A reverse transcription enzyme adds a plurality of non- template oligonucleotides to a nucleotide strand. The reverse transcription reaction can produce single stranded complementary deoxyribonucleic acid (cDNA) molecules each having a molecular tag on a 5’ end thereof, followed by amplification of cDNA to produce a double stranded DNA having the molecular tag on the 5’ end and a 3’ end of the double stranded DNA. [000273] As used herein, the term “wild-type” refers to a gene or gene product that has the characteristics of that gene or gene product when isolated from a naturally occurring source. For example, the amino acid sequence set forth in SEQ ID NO: 7 is a wild-type MMLV amino acid sequence. [000274] In some embodiments of the engineered RT polypeptide or the recombinant RT protein disclosed herein, the RT polypeptide can comprise an amino acid sequence that is at least 90% identical to SEQ ID NO: 1. In some embodiments, the RT polypeptide can comprise an amino acid sequence that is 90-99.99% identical to SEQ ID NO: 1. In some embodiments, the RT polypeptide can comprise an amino acid sequence that is 92-99.99% identical to SEQ ID NO: 1. In some embodiments, the RT polypeptide can comprise an amino acid sequence that is 93-99.99% identical to SEQ ID NO: 1. In some embodiments, the RT polypeptide can comprise an amino acid sequence that is 94-99.99% identical to SEQ ID NO: 1. In some embodiments, the RT polypeptide can comprise an amino acid sequence that is 95-99.99% identical to SEQ ID NO: 1. In some embodiments, the RT polypeptide can comprise an amino acid sequence that is 96-99.99% identical to SEQ ID NO: 1. In some embodiments, the RT polypeptide can comprise an amino acid sequence that is 97-99.99% identical to SEQ ID NO: 1. In some embodiments, the RT polypeptide can comprise an amino acid sequence that is 98-99.99% identical to SEQ ID NO: 1. In some embodiments of the engineered RT polypeptide or the recombinant RT protein disclosed herein, the RT polypeptide can comprise an amino acid sequence that is 90%, 91%, 92%, 93%, 94%, 95%, 96%, 97%, 98%, 99% or 99.5% identical to SEQ ID NO: 1. The amino acid variation are at any one position or combination thereof as identified in an alignment of SEQ ID NO: 1 to any one of the RT polypeptide sequences in Table 1, or Table 2. [000275] In some embodiments, the RT polypeptide can comprise the amino acid sequence set forth in SEQ ID NO: 7. The engineered reverse transcriptase can exhibit an altered reverse 72 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    transcriptase activity as compared to a reverse transcriptase having the amino acid sequence set forth in SEQ ID NO: 1 or 7. [000276] The RT polypeptide can be a variant MMLV reverse-transcriptase having one or more mutations. Specifically, the RT polypeptide contemplated by the present disclosure can comprise a combination of mutations in the amino acid sequence of either the wild-type MMLV (SEQ ID NO 7 or 178) or in a MMLV variant (SEQ ID NO: 1, 143 or 179). [000277] The amino acid sequence of the RT polypeptide sequence contemplated by the present disclosure can be at least 90% identical to SEQ ID NO: 1 or 143. The amino acid sequence of the RT polypeptide sequence can be about 90% to about 99.99% identical to SEQ ID NO: 1 or 143, about 92% to about 99.99% identical to SEQ ID NO: 1 or 143, about 93% to about 99.99% identical to SEQ ID NO: 1 or 143, about 94% to about 99.99% identical to SEQ ID NO: 1 or 143, about 95% to about 99.99% identical to SEQ ID NO: 1 or 143, about 96% to about 99.99% identical to SEQ ID NO: 1 or 143, about 97% to about 99.99% identical to SEQ ID NO: 1 or 143, or about 98% to about 99.99% identical to SEQ ID NO: 1 or 143. In some embodiments, the amino acid sequence of the RT polypeptide sequence can be about 90%, about 91%, about 92%, about 93%, about 94%, about 95%, about 96%, about 97%, about 98%, about 99% or about 99.5% identical to SEQ ID NO: 1 or 143. [000278] As used herein, a “Mutation” refers to a change introduced into a parental or wild type DNA sequence that changes the amino acid sequence encoded by the DNA, including, but not limited to, substitutions, insertions, deletions, point mutations, mutation of multiple nucleotides or amino acids, transposition, inversion, frame shift, nonsense mutations, truncations or other forms of aberration that differentiate the polynucleotide or protein sequence from that of a wild-type sequence of a gene or gene product. The consequences of a mutation include, but are not limited to, the creation of a new character, property, function, or trait not found in the protein encoded by the parental DNA, including, but not limited to, N terminal truncation, C terminal truncation or chemical modification. A “mutation”" also includes an N- or C-terminal extension. In some embodiments, the mutations disclosed herein are substitutions. [000279] In particular, the present disclosure relates to engineered RT polypeptides or recombinant RT polypeptide comprising a wild-type RT or modified reverse transcriptases that comprise one or more (e.g., one, two, three, four, five, ten, twelve, fifteen, twenty, etc.) amino 73 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    acid changes. These amino acid changes render the reverse transcriptase more efficient for nucleic acid synthesis (e.g., single cell profiling assay) requiring very small volume, as compared to an unmutated or an unmodified reverse transcriptase. As will be appreciated by those skilled in the art, one or more of the amino acids identified may be deleted and/or replaced with one or a number of amino acid residues. In a preferred aspect, any one or more of the amino acids may be substituted with any one or more amino acid residues such as Ala, Arg, Asn, Asp, Cys, Gln, Glu, Gly, His, He, Leu, Lys, Met, Phe, Pro, Ser, Thr, Trp, Tyr, and/or Val. [000280] In some embodiments, the RT polypeptide described herein comprises the amino acid sequence of SEQ ID NO:7, and comprises a combination of mutations selected from E69K, L139P, E302R, T306K, W313F, T330P, or N454K; and one or more of M39V, P47L, M66L, F155Y, D200N, D200E, H204R, G429S, L435G, L435K, P448A, D449G, H503V, D524N, T542D, E545G, D583N, H594Q, L603W, L603F, E607K, E607G, P627S, H634Y, H638G, A644V, D653H, K658R or L671P. The engineered polypeptide can comprise the amino acid sequence of SEQ ID NO:7, and comprises a combination of mutations selected from E69K, L139P, D200N, E302R, T306K, W313F, T330P, L435G, P448A, D449G, N454K, D524N, or L603W, and E607K and one or more of M39V, P47L, M66L, F155Y, H204R, G429S, H503V, T542D, E545G, D583N, H594Q, P627S, H634Y, H638G, A644V, D653H, K658R or L671P. [000281] The RT polypeptide sequence can comprise an amino acid sequence that is at least 95% identical to SEQ ID NO:1, 7, or 179, and a combination of mutations indexed to SEQ ID NO:7 or 178 selected from a combination of variants consisting of a T542D mutation, a D583N mutation, an E607G mutation, an A644V mutation, a D653H mutation, and a K658R mutation. The RT polypeptide sequence can comprise an amino acid sequence that is at least 95% identical to SEQ ID NO:1, 7, or 179, and a combination of mutations indexed to SEQ ID NO:7 or 178 selected from a combination of variants consisting of an E545G mutation, a D583N mutation, an H594Q mutation, an L603F mutation, and a S679P mutation. [000282] In some embodiments, the RT polypeptide comprises an amino acid sequence that is at least 90% identical to an amino acid sequence selected from: SEQ ID NO: 14, SEQ ID NO: 22, SEQ ID NO: 23, SEQ ID NO: 24, SEQ ID NO: 25, SEQ ID NO: 26, SEQ ID NO: 27, SEQ ID NO: 28, SEQ ID NO: 29, SEQ ID NO: 30, SEQ ID NO: 31, SEQ ID NO: 32, SEQ ID NO: 33, SEQ ID NO: 35, SEQ ID NO: 36, SEQ ID NO: 37, SEQ ID NO: 38, SEQ ID NO: 39, SEQ 74 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    ID NO: 40, SEQ ID NO: 41, SEQ ID NO: 42, SEQ ID NO: 43, SEQ ID NO: 44, SEQ ID NO: 45, SEQ ID NO: 46, SEQ ID NO: 47, SEQ ID NO: 48, SEQ ID NO: 49, SEQ ID NO: 50, SEQ ID NO: 51, SEQ ID NO: 52, SEQ ID NO: 53, SEQ ID NO: 54, or SEQ ID NO: 55. [000283] In some embodiments, the RT polypeptide can comprise SEQ ID NO: 63, 65, 67, 69, 71, 73, 75, 77, 79, 81, 83, 85, 87, 89, 91, 93, 95, 97, 99, 101, 103, 105, 107, 109, 111, 113, 115, 117, 119, 121, 123, 125, 127, 129, 131, 133, 141, 143, 145, 147, 149, 151, 157, 159, 172, or 173. The RT polypeptide can comprise an amino acid sequence listed in Table 1 or 2. [000284] The amino acid sequence of the RT polypeptide can also comprise E69K, L139P, D200N, E302R, T306K, W313F, T330P, N454K, H503V, D524N, L603W, E607K, and H634Y. In some embodiments, the amino acid sequence of the RT polypeptide comprises a combination of mutations selected from: M66L and L435G; M39V, M66L, and L435K; M39V and L435K; M66L, L435G, P448A and D449G; M39V, M66L, L435G, P448A and D449G; or M66L. [000285] In some embodiments, the amino acid sequence of the RT polypeptide comprises E69K, L139P, D200N, E302R, T306K, W313F, T330P, L435G, P448A, D449G, N454K, D524N, L603W, and E607K; and further comprises a combination of mutations selected from M66L; M66L and H503V; M66L and H634Y; and M66L, H503V, or H634Y. [000286] In some embodiments, the amino acid sequence of the RT polypeptide comprises M39V, E69K, L139P, a D200 mutation, E302R, T306K, W313F, T330P, G429S, P448A, a D449 mutation, L435K, N454K, a L603 mutation, a E607 mutation, and L671P and further comprises a second combination of mutations selected from D524N, T542D, P627S, A644V, D653H, or K658R mutation. In that embodiment, the D200 mutation is a D200N mutation, the D449 mutation is a D449G, the L603 mutation is an L603W, or the E607 mutation is an E607G mutation. [000287] In some embodiments, the amino acid sequence of the RT polypeptide comprises M39V, E69K, L139P, a D200 mutation, E302R, T306K, W313F, T330P, G429S, P448A, a D449 mutation, L435K, N454K, a L603 mutation, a E607 mutation, and L671P and further comprises D524N, T542D, A644V, D653H, an R650H and K658R. In that embodiment, the D200 mutation is a D200N mutation, the D449 mutation is a D449E mutation, the L603 mutation is an L603W mutation, and the E607 mutation is an E607G mutation. 75 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    [000288] In some embodiments, the amino acid sequence of the RT polypeptide comprises M39V, E69K, L139P, a D200 mutation, E302R, T306K, W313F, T330P, G429S, P448A, a D449 mutation, L435K, N454K, a L603 mutation, a E607 mutation, and L671P and further comprises E545G, D583N, and H594Q. In that embodiment, the D200 mutation is a D200N mutation, the D449 mutation is a D449G mutation, the L603 mutation is an L603F mutation, and the E607 mutation is an E607K mutation. [000289] In some embodiments, the amino acid sequence of the RT polypeptide comprises M39V, E69K, L139P, a D200 mutation, E302R, T306K, W313F, T330P, G429S, P448A, a D449 mutation, L435K, N454K, a L603 mutation, a E607 mutation, and L671P and further comprises D524N, T542D, A644V, D653H, and K658R. In that embodiment, the D200 mutation is a D200N mutation, the D449 mutation is a D449E mutation, the L603 mutation is an L603W mutation, and the E607 mutation is an E607G mutation. [000290] In some embodiments, the amino acid sequence of the RT polypeptide comprises M39V, E69K, L139P, a D200 mutation, E302R, T306K, W313F, T330P, G429S, P448A, a D449 mutation, L435K, N454K, a L603 mutation, a E607 mutation, and L671P and further comprises H204R, D524N, T542D, P627S, D583N, A644V, D653H and K658R. In that embodiment, the D200 mutation is a D200E mutation, the D449 mutation is a D449G mutation, the L603 mutation is an L603W mutation, and the E607 mutation is an E607G mutation. [000291] In some embodiments, the amino acid sequence of the RT polypeptide comprises M39V, E69K, L139P, a D200 mutation, E302R, T306K, W313F, T330P, G429S, P448A, a D449 mutation, L435K, N454K, a L603 mutation, a E607 mutation, and L671P and further comprises H204R, E545G, D583N, and H594Q. In that embodiment, the D200 mutation is a D200E mutation, the D449 mutation is a D449G mutation, the L603 mutation is an L603F mutation, and the E607 mutation is an E607K mutation. [000292] In some embodiments, the amino acid sequence of the RT polypeptide comprises M39V, E69K, L139P, a D200 mutation, E302R, T306K, W313F, T330P, G429S, P448A, a D449 mutation, L435K, N454K, a L603 mutation, a E607 mutation, and L671P and further comprises P47L, D524N, T542D, D583N, P627S, A644V, D653H, and K658R. In that embodiment, the D200 mutation is a D200N mutation, the D449 mutation is a D449G mutation, the L603 mutation is an L603W mutation, and the E607 mutation is an E607G mutation. 76 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    [000293] In some embodiments, the RT polypeptide comprises an amino acid sequence that is at least 95% identical to SEQ ID NO: 1 and the amino acid sequence of the engineered reverse transcriptase comprises at least one mutation indexed to SEQ ID NO:7 selected from a M17 mutation; an A32 mutation, a M44 mutation, a M39 mutation, a K47 mutation, a P51 mutation, an M66 mutation, an S67 mutation, an E69 mutation, a L72 mutation, a W94 mutation, a K103 mutation, an R110 mutation, a P117 mutation, an L139 mutation, an F155 mutation, an N178 mutation, an E179 mutation, a T197 mutation, a D200 mutation, an E201 mutation, an H204 mutation, a Q221 mutation, a V223 mutation, a V238 mutation, a G248 mutation, a T265 mutation, an E268 mutation, an R279 mutation, an R280 mutation, a K284 mutation, a T287 mutation, a F291 mutation, an E302 mutation, an E302K mutation, an E302R mutation, a T306 mutation, a T306R mutation, a T306K mutation a P308 mutation, an F309 mutation, a W313 mutation, a T330 mutation, a Y344 mutation, an I347 mutation, a C387 mutation, a W388 mutation, an R389 mutation, a C409 mutation, an R411 mutation, a G413 mutation, an A426 mutation, a G427 mutation, an L435 mutation, an L435G mutation, an L435K mutation, a P448 mutation, a D449 mutation, an R450 mutation, a n N454 mutation, an A480 mutation, an H481 mutation, a N502 mutation, an A502 mutation, an H503 mutation, a D524 mutation, an H572 mutation, a W581 mutation, a D583 mutation, a K585 mutation, an H594 mutation, an L603 mutation, an E607 mutation, an H612 mutation, a P614 mutation, a G615 mutation, an H634 mutation, a P636 mutation, a G637 mutation, an H638 mutation, a D653 mutation, an L671 mutation, or a combination thereof. [000294] In some embodiments, the RT polypeptide comprises an amino acid sequence that is at least 95% identical to SEQ ID NO: 1 and the amino acid sequence of the engineered reverse transcriptase comprises an M39 mutation, a K47 mutation, an L435 mutation, a D449 mutation, a D524 mutation, an E607 mutation, a D653 mutation, and an L671 mutation as indexed to SEQ ID NO:7 and comprising at least one mutation indexed to SEQ ID NO:7 selected from a M17 mutation; an A32 mutation, a M44 mutation, a M39V mutation, a P51 mutation, an M66 mutation, an S67 mutation, an E69 mutation, a L72 mutation, a W94 mutation, a K103 mutation, an R110 mutation, a P117 mutation, an L139 mutation, an F155 mutation, an N178 mutation, an E179 mutation, a T197 mutation, a D200 mutation, an E201 mutation, an H204 mutation, a Q221 mutation, a V223 mutation, a V238 mutation, a G248 mutation, a T265 mutation, an E268 mutation, an R279 mutation, an R280 mutation, a K284 mutation, a T287 mutation, a F291 77 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    mutation, an E302 mutation, an E302K mutation, an E302R mutation, a T306 mutation, a T306R mutation, a T306K mutation a P308 mutation, an F309 mutation, a W313 mutation, a T330 mutation, a Y344 mutation, an I347 mutation, a C387 mutation, a W388 mutation, an R389 mutation, a C409 mutation, an R411 mutation, a G413 mutation, an A426 mutation, a G427 mutation, an L435G mutation, an L435K mutation, a P448 mutation, a D449G mutation, an R450 mutation, a n N454 mutation, an A480 mutation, an H481 mutation, a N502 mutation, an A502 mutation, an H503 mutation, a D524N mutation, an H572 mutation, a W581 mutation, a D583 mutation, a K585 mutation, an H594 mutation, an L603 mutation, an H612 mutation, a P614 mutation, a G615 mutation, an H634 mutation, a P636 mutation, a G637 mutation, an H638 mutation, or a combination thereof. [000295] In other embodiments, the engineered RT polypeptide exhibits an altered reverse transcriptase related activity when compared to a reverse transcriptase having the amino acid sequence set forth in SEQ ID NO: 1. [000296] In some embodiments, an RT polypeptide comprises an amino acid sequence that is at least 95% identical to SEQ ID NO: 1. In other embodiments, the engineered reverse transcriptase exhibits an altered reverse transcriptase related activity as compared to a reverse transcriptase having the amino acid sequence set forth in SEQ ID NO: 1. In additional embodiments, the RT polypeptide comprises a combination of mutations indexed to SEQ ID NO:7 selected from: (i) an E69K mutation, an E302R mutation, a T306K mutation, a W313F mutation, a L435G mutation, or an N454K mutation, and comprising at least one mutation selected from an M39V mutation, an M66L mutation, an L139P mutation, an F155Y mutation, a D200N mutation, an E201Q mutation, a T287A mutation, a T330P mutation, an R411F mutation, a P448A mutation, a D449G mutation, an H503V mutation, an H594K mutation, L603W mutation, an E607K mutation, an H634Y mutation, a G637R mutation and an H638G mutation; (ii) an L139P mutation, a D200N mutation, a T330P mutation, an L603W mutation, or an E607K mutation, and comprising at least one mutation selected from: an M39V mutation, an M66L mutation an E69K mutation, an F155Y mutation, an E201Q mutation, a T287A mutation, an E302R mutation, a T306K mutation, a W313F mutation, an R411F mutation, an L435G mutation, a P448A mutation, a D449G mutation, an N454K mutation, an H503V mutation, an H594K mutation, an H634Y mutation, a G637R mutation or an H638G mutation; (iii) an A32V 78 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    mutation, an L72R mutation, a D200C mutation, a G248C mutation, an E286R mutation, an E302R mutation, a W388R mutation, and an L435G mutation; or (iv) a Y344L mutation and an I347L mutation. [000297] In some embodiments, the RT polypeptide comprises an amino acid sequence that is at least 95% identical to a reverse transcriptase having the amino acid sequence set forth in SEQ ID NO:1. In some embodiments, the engineered reverse transcription enzyme comprises an amino acid sequence that is at least 95% identical to SEQ ID NO: 1 and has at least one mutation selected from the group consisting of an M39V mutation, a P47L mutation, M66L mutation, an E69K mutation, an L139P mutation, a D200N mutation, an H204R mutation, an E302R mutation, a T306K mutation, a W313F mutation, a T330P mutation, an L435G mutation, a G429S mutation, an L435K mutation, a P448A mutation, a D449G mutation, a N454K mutation, an H503V mutation, a D524N mutation, a T542 mutation, an E545G mutation, a D583N mutation, an H594Q mutation, an L603W mutation, an E607K mutation, a P627S mutation, an H634Y mutation, an A644V mutation, an R650H mutation, a D653H mutation, a K658R mutation, an L671P mutation, or an S679P mutation; and the engineered RT polypeptide or recombinant RT protein described herein exhibits an altered reverse transcriptase related activity. [000298] In some embodiments, the disclosure provides an engineered RT polypeptide comprising an amino acid sequence that is at least 95% identical to SEQ ID NO:1 and the amino acid sequence of the engineered reverse transcriptase comprises a combination of mutations indexed to SEQ ID NO:7 or 178 selected from the group consisting of (a) an E69K mutation, an L139P mutation, a D200N mutation, an E302R mutation, a T306K mutation, a W313F mutation, a T330P mutation, a N454K mutation, an H503V mutation, a D524N mutation, an L603W mutation, an E607K mutation, and an H634Y mutation; (b) an M66L mutation, an E69K mutation, an L139P mutation, a D200N mutation, an E302R mutation, a T306K mutation, a W313F mutation, a T330P mutation, a N454K mutation, a D524N mutation, an H503V mutation, an L603W mutation, an E607K mutation, and an H634Y mutation, and at least one mutation selected from the group consisting of an L435G mutation, an L435K mutation, an M39V mutation, a P448A mutation and a D449G mutation; (c) an M39V mutation, an E69K mutation, an L139P mutation, a D200N mutation, an E302R mutation, a T306K mutation, a W313F mutation, a T330P mutation, a N454K mutation, an H503V mutation, a D524N 79 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    mutation, an L603W mutation, an E607K mutation, and an H634Y mutation; and (d) an M39V mutation, an E69K mutation, an L139P mutation, a D200 mutation, an E302R mutation, a T306K mutation, a W313F mutation, a T330P mutation, an L435K mutation, a G429S mutation, a P448A mutation, a D449 mutation, an N454K mutation , an L603 mutation, an E607 mutation and an L671P mutation. In that embodiment, the D200 mutation is selected from the group consisting of D200N and D200E. In that embodiment, the D449 mutation is selected from the group consisting of D449G an D449E. In that embodiment, the L603 mutation is selected from the group consisting of L603W and L603F. In that embodiment, the E607 mutation is selected from the group consisting of E607G and E607K. In another embodiment, the engineered RT polypeptide further comprises at least one mutation selected from the group consisting of P47L, H204R, D524N, T542D, E545G, D583N, H594Q, P627S, A644V, R650H, D653H, K658R, L671P, and S679P. [000299] In some embodiments, an engineered reverse transcriptase of the present application has an amino acid sequence that is at least 95% identical to SEQ ID NO:1 and the amino acid sequence of said engineered reverse transcriptase comprises a combination of mutations indexed to SEQ ID NO:7 or 178; and the amino acid sequence of said engineered reverse transcriptase comprises a combination of mutations selected from the group consisting of: an E69K mutation, an L139P mutation, a D200N mutation, an E302R mutation, a T306K mutation, a W313F mutation, a T330P mutation, a N454K mutation, an H503V mutation, a D524N mutation, an L603W mutation, an E607K mutation, and an H634Y mutation and further comprising a second combination of mutations selected from the group consisting of: (a) an M66L mutation and an L534G mutation, (b) an M39V mutation, an M66L mutation and an L435K mutation, (c) an M39V mutation and an L435K mutation, (d) an M66L mutation, an L435G mutation, a P448 mutation, and D449G mutation, and (e) an M39V mutation, an M66L mutation, an L435G mutation, a P448 mutation and a D449G mutation. [000300] In some embodiments, an engineered reverse transcriptase of the present application has an amino acid sequence that is at least 95% identical to SEQ ID NO:1 and the amino acid sequence of the engineered reverse transcriptase comprises a combination of mutations selected from the group consisting of: an M39V mutation, an E69K mutation, an L139P mutation, a D200 mutation, an E302R mutation, a T306K mutation, a W313F mutation, a T330P mutation, a 80 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    G429S mutation a P448A mutation, a D449 mutation, an L435K mutation, a N454K mutation, an L603 mutation, an E607 mutation, and an L671P mutation and further comprising a second combination of mutations selected from the group consisting of: (a) a D524N mutation, a T542D mutation, an A644V mutation, a D653H mutation, a K658R mutation, a S679P mutation, and wherein said D200 mutation is a D200N mutation, said D449 mutation is a D449G, said L603 mutation is an L603W, said E607 mutation is an E607G mutation, and a P627S mutation; (b) a D524N mutation, a T542D mutation, an A644V mutation, a D653H mutation, an R650 mutation and a K658R mutation, and wherein said D200 mutation is a D200N mutation, said D449 mutation is a D449E mutation, said L603 mutation is an L603W mutation, and said E607 mutation is an E607G mutation; (c) an E545G mutation, a D583N mutation, an H594Q mutation, and an S679P mutation, and wherein said D200 mutation is a D200N mutation, said D449 mutation is a D449G mutation, said L603 mutation is an L603F mutation, said E607 mutation is an E607K mutation; (d) a D524N mutation, a T542D mutation, an A644V mutation, a D653H mutation, and a K658R mutation, and wherein said D200 mutation is a D200N mutation, said D449 mutation is a D449E mutation, said L603 mutation is an L603W mutation, said E607 mutation is an E607G mutation; (e) an H204R mutation, a D524N mutation, a T542D mutation, a D583N mutation, an A644V mutation, a D653H mutation, and a K658R mutation, wherein said D200 mutation is a D200E mutation, said D449 mutation is a D449G mutation, said L603 mutation is an L603W mutation, said E607 mutation is an E607G mutation, and a P627S mutation, (f) an H204R mutation, an E454G mutation, a D583N mutation, an H594Q mutation, and an S679P mutation, wherein said D200 mutation is a D200E mutation, said D449 mutation is a D449G mutation, said L603 mutation is an L603F mutation, said E607 mutation is an E607K mutation; and (g) a P47 mutation, a D524N mutation, a T542D mutation, a D583N mutation, an A644V mutation, a D653H mutation, a K658R mutation and an S679P mutation. In that embodiment, the P47 mutation is a P47L mutation, the D200 mutation is a D200N mutation, the D449 mutation is a D449G mutation, the L603 mutation is an L603W mutation, the E607 mutation is an E607G mutation, and the P627 mutation is a P627S mutation. [000301] A variant may comprise a first combination of mutations or alterations and may comprise an additional or second combination of mutations. 81 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    [000302] A first combination of mutations or alterations may include, but is not limited to, a combination of: (1) a M39 mutation, a K47 mutation, an L435 mutation, a D449 mutation, a D524 mutation, an E607 mutation, a D653 mutation and an L671 mutation; (2) an M39V mutation, a K47 mutation, an L435K mutation, a D449G mutation, a D524N mutation, an E607 mutation, a D653 mutation and an L671 mutation; (3) an M39 mutation, an M66 mutation, an E302 mutation, a T306 mutation, an L435 mutation, a D449 mutation, a D524 mutation, an E607 mutation, a D653 mutation and an L671 mutation; (4) an M39 mutation, an M66 mutation, an E302 (K or R) mutation, a T306 (R or K) mutation, an L435 (K or G), a D449 mutation, a D524 mutation, an E607 (G or K) mutation, a D653 mutation, and an L671 mutation; and (5) an M39V mutation, an M66 mutation, an E302 (K or R) mutation, a T306 (R or K) mutation, an L435 (K or G), a D449G mutation, a D524N mutation, an E607 (G or K) mutation, a D653 mutation, and an L671 mutation. [000303] The second combination of mutations in a first engineered reverse transcriptase may comprise either a different set of mutations or a partially different second set of mutations as in a second engineered reverse transcriptase. A second combination of mutations or alterations may include but is not limited to: (a) one or more mutations selected from an M17 mutation; an A32 mutation, a M44 mutation, a P51 mutation, an M66 mutation, an S67 mutation, an E69 mutation, a L72 mutation, a W94 mutation, a K103 mutation, an R110 mutation, a P117 mutation, an L139 mutation, an F155 mutation, an N178 mutation, an E179 mutation, a T197 mutation, a D200 mutation, an E201 mutation, an H204 mutation, a Q221 mutation, a V223 mutation, a V238 mutation, a G248 mutation, a T265 mutation, an E268 mutation, an R279 mutation, an R280 mutation, a K284 mutation, a T287 mutation, a F291 mutation, an E302 mutation, an E302K mutation, an E302R mutation, a T306 mutation, a T306R mutation, a T306K mutation, a P308 mutation, an F309 mutation, a W313 mutation, a T330 mutation, a Y344 mutation, an I347 mutation, a C387 mutation, a W388 mutation, an R389 mutation, a C409 mutation, an R411 mutation, a G413 mutation, an A426 mutation, a G427 mutation, an L435G mutation, an L435K mutation, a P448 mutation, a D449G mutation, an R450 mutation, an N454 mutation, an A480 mutation, an H481 mutation, a N502 mutation, an A502 mutation, an H503 mutation, a D524N mutation, an H572 mutation, a W581 mutation, a D583 mutation, a K585 mutation, an H594 mutation, an L603 mutation, an H612 mutation, a P614 mutation, a G615 mutation, an H634 mutation, a P636 mutation, a G637 mutation, or an H638 mutation; (b) an E69K mutation, an 82 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    E302R mutation, a T306K mutation, a W313F mutation, an L435G mutation, and an N454K mutation, and comprising at least one mutation selected from the group consisting of an M39V mutation, an M66L mutation, an L139P mutation, an F155Y mutation, a D200N mutation, an E201Q mutation, a T287A mutation, a T330P mutation, an R411F mutation, a P448A mutation, a D449G mutation, an H503V mutation, an H594K mutation, L603W mutation, an E607K mutation, an H634Y mutation, a G637R mutation and an H638G mutation; (c) an L139P mutation, a D200N mutation, a T330P mutation, an L603W mutation, and an E607K mutation, and comprising at least one mutation selected from the group consisting of: an M39V mutation, an M66L mutation, an E69K mutation, an F155Y mutation, an E201Q mutation, a T287A mutation, an E302R mutation, a T306K mutation, a W313F mutation, an R411F mutation, an L435G mutation, a P448A mutation, a D449G mutation, an N454K mutation, an H503V mutation, an H594K mutation, an H634Y mutation, a G637R mutation and an H638G mutation; (d) an A32V mutation, an L72R mutation, a D200C mutation, a G248C mutation, an E286R mutation, an E302R mutation, a W388R mutation, or an L435G mutation; and (e) a Y344L mutation and an I347L mutation. It is recognized that the second combination of mutations may comprise a group of mutations as described herein and one or more additional mutations. [000304] In non-limiting embodiments, the engineered RT variants of the present disclosure comprise a M39V, M66I, Q91R, I347V, H594Q, or a combination thereof in the RT backbone of SEQ ID NO: 143 or SEQ ID NO: 7. In non-limiting embodiments, the engineered RT polypeptide comprises: M39V, M66I, Q91R, I347V, and H594Q (SEQ ID NO: 129 , SOLD 034). In non-limiting embodiments, the engineered RT variants of the present disclosure comprise M39V, T542D, D583N, E607G, A644V, D653H, K658R, L671P, or a combination thereof in the RT backbone of SEQ ID NO: 143 or SEQ ID NO: 7. In non-limiting embodiments, the engineered RT polypeptide comprises: M39V, T542D, D583N, E607G, A644V, D653H, K658R, and L671P (SEQ ID NO: 111, SOLD 025). [000305] In some embodiments, the engineered RT polypeptide comprises: M39V, M66I, Q91R, I347V, H594Q, or a combination thereof, and optionally M39V, M66I, Q91R, I347V, H594Q, or the combination thereof (substituted) in the RT sequence of SEQ ID NO: 143 (SEQ ID NO: 129, SOLD 034). 83 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    [000306] In some embodiments, the engineered RT polypeptide comprises: M39V, T542D, D583N, E607G, A644V, D653H, K658R, L671P, or a combination thereof, and optionally M39V, T542D, D583N, E607G, A644V, D653H, K658R, L671P, or the combination thereof substituted in the RT sequence of SEQ ID NO: 143 (or SEQ ID NO: 7) (SEQ ID NO: 111, SOLD 025). In some embodiments, the engineered RT polypeptide comprises SOLD 33 VDG or SEQ ID NO: 173. [000307] In some embodiments, the engineered RT polypeptide described herein comprises an amino acid sequence that is at least about 90% identical to an amino acid sequence selected from the group consisting of SEQ ID NO: 27-61, 63, 65, 67, 69, 71, 73, 75, 77, 79, 81, 83, 85, 87, 89, 91, 93, 95, 97, 99, 101, 103, 105, 107, 109, 111, 113, 115, 117, 119, 121, 123, 125, 127, 129, 131, 133, 141, 143, 145, 147, 149, 151, 157, 159, 172, and 173. [000308] In some embodiments of the engineered RT polypeptide described herein, the RT polypeptide is 42B L (SEQ ID NO: 145), 50A+G (SEQ ID NO: 147), SOLD 022 (SEQ ID NO: 105), SOLD 023 (SEQ ID NO: 107), SOLD 025 (SEQ ID NO: 111), SOLD 031 (SEQ ID NO: 123), SOLD 033 (SEQ ID NO: 127), SOLD 034 (SEQ ID NO: 129), SOLD 035 (SEQ ID NO: 131), SOLD 001 (SEQ ID NO: 65), SOLD 33 VDG (SEQ ID NO: 173), or an RT polypeptide set forth in SEQ ID NO: 143, or SEQ ID NO: 172. E. Engineered DAT1 Reverse Transcriptases [000309] One aspect of the present disclosure provides an engineered RT polypeptide comprising an amino acid sequence that is at least 90%, at least 92%, at least 95%, at least 97%, at least 98%, or at least 99% identical to an amino acid sequence of an RT disclosed in Table 1, or Table 2, and a DNA binding domain comprising an amino acid sequence selected from the group consisting of SEQ ID NO: 2, 3, 5, 6, 8, and 9. [000310] Another aspect of the present disclosure provides an engineered RT polypeptide comprising an amino acid sequence that is at least 90%, at least 92%, at least 95%, at least 97%, at least 98%, or at least 99% identical to an amino acid sequence of SEQ ID NOs: 27-61, 63, 65, 67, 69, 71, 73, 75, 77, 79, 81, 83, 85, 87, 89, 91, 93, 95, 97, 99, 101, 103, 105, 107, 109, 111, 113, 115, 117, 119, 121, 123, 125, 127, 129, 131, 133, 141, 143, 145, 147, 149, 151, 157, 159, 172, or 173; and a DNA binding domain comprising an amino acid sequence selected from the group consisting of SEQ ID NO: 2, 3, 5, 6, 8, and 9. 84 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    [000311] Another aspect of the present disclosure provides an engineered RT polypeptide comprising an amino acid sequence that is at least 90%, at least 92%, at least 95%, at least 97%, at least 98%, or at least 99% identical to an amino acid sequence of SEQ ID NOs: 27-61, 63, 65, 67, 69, 71, 73, 75, 77, 79, 81, 83, 85, 87, 89, 91, 93, 95, 97, 99, 101, 103, 105, 107, 109, 111, 113, 115, 117, 119, 121, 123, 125, 127, 129, 131, 133, 141, 143, 145, 147, 149, 151, 157, 159, 172, or 173; and a DNA binding domain comprising an amino acid sequence selected from the group consisting of SEQ ID NO: 11-24. [000312] Another aspect of the present disclosure provides an engineered RT polypeptide comprising an amino acid sequence of an RT disclosed in Table 1 or Table 2; and an amino acid sequence of DNA binding domain disclosed in Table 1. [000313] In some embodiments, the engineered RT polypeptide described herein comprises the amino acid sequence of any one of SEQ ID NO: 174-188. [000314] The engineered RT polypeptide described herein can comprise 42B L RT (SEQ ID NO: 145) operably linked to a full length DAT1 at the N-terminus. In that embodiment, the engineered RT can comprise the amino acid sequence of SEQ ID NO: 183. The engineered RT polypeptide described herein can comprise 42B L RT operably linked to a full length DAT1 at the C-terminus. In that embodiment, the engineered RT can comprise the amino acid sequence of SEQ ID NO: 184. [000315] DAT1(90) can be fused to any reverse transcriptase enzymes described herein, such as variants of MMLV reverse transcriptases disclosed in Table 1 or Table 2. While the majority of engineered RT polypeptide embodiments disclosed herein are N-terminus fusion proteins, the present disclosure also contemplates engineered RT polypeptides where the DNA binding domain, e.g., DAT1 or variant thereof is operably linked to the C-terminus of any RT polypeptide described herein. Accordingly, the present disclosure provides any combination of DNA binding domain and RT polypeptide described in Table 1 or Table 2. Non-exhaustive list of possible engineered RT polypeptides or recombinant RT proteins can comprise reverse transcriptases fused to full length DAT protein, DAT1 truncated to 36 amino acids, or DAT1(90) homologs. 85 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    [000316] For example, the present disclosure provides an engineered RT comprising the amino acid sequence of SEQ ID NO: 178 (e.g., N-DAT-9042B). The engineered RT can comprise the amino acid sequence of SEQ ID NO: 179 (e.g., N-DAT-9050A+ G). The engineered RT can comprise the amino acid sequence of SEQ ID NO: 180 (e.g., N-DAT-90 SOLD 33 VDG). The engineered RT can comprise the amino acid sequence of SEQ ID NO: 176 (e.g., N-DAT-90 SOLD 01). The engineered RT can comprise the amino acid sequence of SEQ ID NO: 177 (e.g., C-DAT-90 SOLD 01). As disclosed in FIGs.20-33, an engineered RT polypeptide comprising N-DAT-9042B (SEQ ID NO: 178) and/or N-DAT-9050A+ G (SEQ ID NO: 178) were shown to enhance the sensitivity of the 5’ single cell assay as shown by enhanced transcript capture. [000317] In some embodiments, the engineered RT polypeptide described herein comprises 42B L RT (SEQ ID NO: 145) operably linked to an N-terminal truncated variant of DAT1 (D90) comprising the first 90 amino acids of the full length DAT1 at the C-terminus. In that embodiment, the engineered RT can comprise the amino acid sequence of SEQ ID NO: 174. The engineered RT polypeptide described herein can comprise 42B L RT operably linked to an N- terminal truncated variant of DAT1 (D90) comprising the first 90 amino acids of the full length DAT1 at the N-terminus. In that embodiment, the engineered RT can comprise the amino acid sequence of SEQ ID NO: 175. [000318] The engineered RT polypeptide described herein can comprise 42B RT operably linked to an N-terminal truncated variant of DAT1 (D90) comprising the first 90 amino acids of the full length DAT1 at the N-terminus. In that embodiment, the engineered RT can comprise the amino acid sequence of SEQ ID NO: 178. [000319] The engineered RT polypeptide described herein can comprise SOLD 01 RT operably linked to an N-terminal truncated variant of DAT1 (D90) comprising the first 90 amino acids of the full length DAT1 at the N-terminus. In that embodiment, the engineered RT can comprise the amino acid sequence of SEQ ID NO: 176. The engineered RT polypeptide described herein can comprise SOLD 01 operably linked to an N-terminal truncated variant of DAT1 (D90) comprising the first 90 amino acids of the full length DAT1 at the C-terminus. In that embodiment, the engineered RT can comprise the amino acid sequence of SEQ ID NO: 177. [000320] The engineered RT polypeptide described herein can comprise 50A+G RT operably linked to an N-terminal truncated variant of DAT1 (D90) comprising the first 90 amino acids of 86 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    the full length DAT1 at the N-terminus. In that embodiment, the engineered RT can comprise the amino acid sequence of SEQ ID NO: 179. [000321] The engineered RT polypeptide described herein can comprise SOLD 33 VDG RT operably linked to an N-terminal truncated variant of DAT1 (D90) comprising the first 90 amino acids of the full length DAT1 at the N-terminus. In that embodiment, the engineered RT can comprise the amino acid sequence of SEQ ID NO: 180. [000322] The DAT1 protein can be truncated to a minimal binding domain of 36 amino acids. In such embodiments, the engineered RT polypeptide described herein can comprise 42B L RT (SEQ ID NO: 145) operably linked to an N-terminal truncated variant of DAT1 (D36) comprising the first 36 amino acids of the full length DAT1 at the N-terminus (e.g., N- Dat36_42BL). In that embodiment, the engineered RT can comprise the amino acid sequence of SEQ ID NO: 181. [000323] The engineered RT polypeptide described herein can comprise 42B L RT (SEQ ID NO: 145) operably linked to an N-terminal truncated variant of DAT1 (D36) comprising the first 36 amino acids of the full length DAT1 at the C-terminus (e.g., C-DAT36_42BL). In that embodiment, the engineered RT can comprise the amino acid sequence of SEQ ID NO: 182. [000324] Homologs of DAT1 from non-yeast organisms can also be used to engineer the engineered the RT polypeptide or recombinant RT protein disclosed herein. Exemplary embodiments of such engineered RT polypeptides include, but not limited to an engineered RT polypeptide comprising the amino acid sequence of SEQ ID NO: 185 (e.g., N-DAT1-TL- QID04042BL). The engineered RT polypeptide can comprise the amino acid sequence of SEQ ID NO: 186 (e.g., N-DAT1-TL-XP36142BL). The engineered RT polypeptide can comprise the amino acid sequence of SEQ ID NO: 187 (e.g., N-DAT1-TL-XP55842BL). The engineered RT polypeptide can comprise the amino acid sequence of SEQ ID NO: 188 (e.g., N-DAT1-TL- XP68342BL). In some embodiments, SEQ ID NO: 185-188 can comprise a DAT1 protein from different budding yeasts. [000325] While embodiments disclosed herein only provide N-terminal fusion protein. The present disclosure contemplates engineered RT proteins comprising DAT1 from other organisms where the DAT1 RT proteins are C-terminal fusions. In some embodiments of the engineered RT 87 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    polypeptide described herein, the DNA binding domain (e.g., DAT1 protein or variant thereof) can contain at least three repeated pentads of G-R-K-P-G. [000326] In some embodiments, the engineered RT polypeptide described herein comprises an amino acid sequence having at least 90%, at least 91%, at least 92%, at least 93%, at least 94%, at least 95%, at least 96%, at least 97%, at least 98%, or at least 99% sequence identity to any one of SEQ ID NO: SEQ ID NO: 174-188. In some embodiments, the engineered RT polypeptide described herein comprises an amino acid sequence having 90%, 91%, 92%, 93%, 94%, 95%, 96%, 97%, 98%, or 99% sequence identity to any one of SEQ ID NO: SEQ ID NO: 174-188. [000327] In some embodiments, the engineered RT disclosed herein further comprises a Tag protein as disclosed herein. The tag protein can be selected from the group consisting of an affinity tag, a fluorescent tag, or an expression and/or solubility enhancement tag. the tag is selected from hexahistidine tag (his-tag), small ubiquitin-like modifier tag (SUMO), a VariFlex C-Terminal solubility enhancement tag, a short peptide C-terminal tag, Thioredoxin (Trx) tag, Solubility-enhancer peptide sequences (SET) tag, IgG domain B1 of Protein G (GB1) tag, IgG repeat domain ZZ of Protein A (ZZ) tag, Solubility enhancing Ubiquitous Tag (SNUT tag), Seventeen kilodalton protein (Skp tag), Phage T7 protein kinase (T7PK) tag, E. coli secreted protein A (EspA) tag, Monomeric bacteriophage T70.3 protein (Orc protein) (Mocr) tag, E. coli trypsin inhibitor (Ecotin) tag, Calcium-binding protein (CaBP) tag, Stress-responsive arsenate reductase (ArsC) tag, N-terminal fragment of translation initiation factor IF2 (IF2-domain I) tag, N-terminal fragment of translation initiation factor IF2 (Expressivity) tag, Fasciola hepatica 8- kDa antigen tag (Fh8), Glutathione-S-transferase (GST) tag, maltose-binding protein tag (MBP), Flag tag peptide (FLAG), streptavidin binding peptide tag (Strep-II; strep), calmodulin-binding protein tag (CBP), mutated dehalogenase tag (HaloTag), staphylococcal Protein A (Protein A), intein mediated purification with the chitin-binding domain (IMPACT (CBD)), cellulose-binding module (CBM), dockerin domain of Clostridium josui tag (Dock), or fungal avidin-like protein (Tamavidin) [000328] The engineered RT polypeptide or the recombinant RT protein can comprise an hexahistidine tag (his-tag). Alternatively, the engineered RT polypeptide or the recombinant RT 88 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    protein can comprise an amino acid sequence of SEQ ID NO: 62; or an amino acid sequence having at least 90% sequence identity to the amino acid sequence of SEQ ID NO: 62. F. Enhanced Reverse Transcriptase activity [000329] The engineered reverse transcriptase of the present disclosure can be a variant MMLV reverse-transcriptase with increased or enhanced reverse transcriptase activity. The term “increased” reverse transcriptase activity refers to the level of reverse transcriptase activity of a variant (e.g., mutant reverse transcriptase enzyme (e.g., MMLV variants disclosed herein)) as compared to its wild-type form (e.g., WT MMLV or MMLV having the amino acid of SEQ ID NO: 7) or a known variant (e.g., MMLV having the amino acid of SEQ ID NO: 1). A mutant enzyme is said to have an "increased" reverse transcriptase activity if the level of its reverse transcriptase activity (as measured by methods described herein or known in the art) is at least 10% or more than its wild-type or a known variant. For example, the variant can have at least 10%, 15%, 20%, 25%, 30%, 40%, 50%, 60%, 70%, 80%, 90%, 100% more or at least 2-fold, 3- fold, 4-fold, 5-fold, or 10-fold or more activity than the wild-type or known variant. [000330] An engineered reverse transcriptase may exhibit one or more reverse transcriptase related activities including but not limited to, an RNA-dependent DNA polymerase activity, an RNAse H activity, a DNA-dependent DNA polymerase activity, an RNA binding activity, a DNA binding activity, a polymerase activity, a primer extension activity, a strand-displacement activity, a helicase activity, a strand transfer activity, a template binding activity, transcription template switching, transcription efficiencies, template switching efficiencies, processivity efficiencies, incorporation efficiencies, fidelity efficiencies, polymerization efficiencies, altered specificity, altered non-templated base addition, altered thermostability, altered tailing, altered adapter binding, binding efficiencies, ability to yield unique molecular identifiers (UMI), ability to yield median UMI, transcription efficiency, template switching efficiency, processivity, incorporation efficiency, Kd, distribution, fidelity, polymerization efficiency, Km, specificity, non-templated base addition, thermostability, tailing, adapter binding, binding efficiency, binding affinity (Km/Kcat), Vmax and ability to yield median UMI/cell and altered binding affinities. [000331] A change in any activity may increase, decrease or have no effect on a different reverse-transcriptase related activity. In addition, a change in one activity may alter multiple 89 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    properties of a reverse transcriptase. When multiple properties are affected, the properties may be altered similarly or differently. Methods of evaluating reverse transcriptase related activities are known in the art. A change in a reverse transcriptase related activity may alter one or more of the following results including but not limited to the yield of unique molecular identifiers (UMI), the median UMI obtained, the yield of mitochondrial UMI counts, and/or the yield of ribosomal UMI counts. A change or alteration in the yield of UMI the median UMI obtained, the yield of mitochondrial UMI counts, and/or the yield of ribosomal UMI counts may indicate one or more altered reverse transcriptase related activities. [000332] The engineered reverse transcription enzyme variants of the present disclosure unexpectedly provide an altered or improved reverse transcriptase activity, such as but not limited to, improved template switching (TS) efficiency, higher end-to-end template jumping/switching, improved processivity efficiency, improved binding affinity, improved transcription efficiency, improved chemical tolerance, improved ability to yield mitochondrial unique molecular identifier (UMI) counts, improved ability to yield ribosomal unique molecular identifier (UMI) counts, improved shelf life, higher strand displacement, increased thermostability, improved thermoreactivity, and any combination thereof. An engineered reverse transcription enzyme of the current application may exhibit an altered base-biased template switching activity such as an increased base-biased template switching activity, decreased base- biased template switching activity or an altered base-bias to the template switching activity. [000333] In some embodiments of the engineered RT polypeptide described herein, or the recombinant RT protein described herein, the engineered RT polypeptide or the recombinant RT protein exhibits increased template switching (TS) efficiency, increased processivity efficiency, increased binding affinity, increased transcription efficiency, increased chemical tolerance, improved ability to yield mitochondrial unique molecular identity (UMI) counts, improved ability to yield ribosomal unique molecular identity (UMI) counts, longer shelf life, higher strand displacement, higher end-to-end template jumping, or any combination thereof, when compared to an RT polypeptide or a recombinant RT protein lacking a conjugated DNA binding domain. [000334] In some embodiments of the engineered RT polypeptide described herein, or the recombinant RT protein described herein, the engineered RT polypeptide or the recombinant RT protein comprises at least two or more of increased template switching (TS) efficiency, increased 90 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    processivity efficiency, increased binding affinity, increased transcription efficiency, increased chemical tolerance, improved ability to yield mitochondrial unique molecular identity (UMI) counts, longer shelf life, higher strand displacement, higher end-to-end template jumping, or improved ability to yield ribosomal unique molecular identity (UMI) counts, when compared to an RT polypeptide or a recombinant RT protein lacking a conjugated DNA binding domain. [000335] In some embodiments of the engineered RT polypeptide described herein or the recombinant RT protein described herein, the recombinant RT protein or the engineered RT exhibits increased transcript capture during amplification, when compared to an RT polypeptide or a recombinant RT protein lacking a conjugated DNA binding domain. [000336] In some embodiments of the engineered RT polypeptide or the recombinant RT protein described herein, the DNA binding domain enhances the hybridization of a transcript and a primer during a nucleic acid amplification process, when compared to an RT polypeptide or a recombinant RT protein lacking a conjugated DNA binding domain. [000337] The engineered reverse transcription enzyme variants of the present disclosure unexpectedly provided an altered reverse transcriptase activity, such as but not limited to, improved thermal stability, processive reverse transcription, non-templated base addition, binding affinity, and template switching ability. An engineered reverse transcription enzyme of the current application may exhibit an altered base-biased template switching activity such as an increased base-biased template switching activity, decreased base-biased template switching activity or an altered base-bias to the template switching activity. An engineered reverse transcriptase variant may exhibit enhanced template switching with a 5’-G cap on the substrate. Furthermore, an engineered reverse transcription enzyme variants described herein may also exhibit unexpectedly higher resistance to cell lysate (i.e., are less inhibited by cell lysate) than that exhibited by an enzyme having the amino acid sequence set forth in SEQ ID NO:1. Lastly, an engineered reverse transcription enzyme variants of the present disclosure may have an unexpectedly greater ability to capture full-length transcripts (e.g., in T-cell receptor paired transcriptional profiling), as compared to that exhibited by an enzyme having the amino acid sequence set forth in SEQ ID NO:1. [000338] It is recognized that mutation of one or more residues may alter a first reverse transcriptase activity differently than a second reverse transcriptase activity. Further it is 91 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    recognized that a different combination of mutations, such as different sites or residue changes may alter a reverse transcriptase activity similarly or differently. The variants that can template switch in the 5’ assay share the following alterations: E69K, E302R, T306K, W313F, L/K435G, and N454K. These variants may further comprise additional alterations that may affect one or more reverse transcriptase related activities. M39V and M66L may improve template switching. Without being limited by mechanism, variants comprising a M39V or a M66L mutation that do not exhibit altered performance in the 5’ GEM assay may exhibit an altered processivity, an altered kd or both. L435K mutants may improve thermostability in the presence of primer template. In the absence of primer template L435K variants may exhibit a thermal denaturation profile similar to that of the wild-type protein. L435K, P448 and D449 are residues in the connection domain; altering these residues may result in increased conformational flexibility. Additionally, the connection domain is thought to impact the conformational flexibility of the RNAse H domain. H503 and H634 occur within the RNAse H domain. The H503V and H634Y variants may impact primer-template contacting, processivity or both primer-template contacting and processivity. [000339] Some variants share the following alterations: (a) the combination of variants consisting of a T542D mutation, a D583N mutation, an E607G mutation, an A644V mutation, a D653H mutation, and a K658R mutation. Some variants share the following alterations: (b) the combination of variants consisting of an E545G mutation, a D583N mutation, an H594Q mutation, an L603F mutation, and a S679P mutation. These variants may further comprise additional alterations that may affect one or more reverse transcriptase related activities. The combination of variants consisting of a T542D mutation, a D583N mutation, an E607G mutation, an A644V mutation, a D653H mutation, and a K658R mutation and the combination of variants consisting of an E545G mutation, a D583N mutation, an H594Q mutation, an L603F mutation and a S679P mutation may exhibit an altered RNAse H activity. 1. RNase H activity [000340] In some embodiments, the engineered reverse transcriptase polypeptide or recombinant RT protein is engineered to have reduced and/or abolished RNase activity. RNase H activity refers to endoribonuclease degradation of the RNA of a DNA-RNA hybrid to produce 5' phosphate terminated oligonucleotides that are 2-9 bases in length. RNase H activity does not 92 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    include degradation of single-stranded nucleic acids, duplex DNA, or double-stranded RNA. Removal of the RNase H activity of reverse transcriptase can eliminate the problem of RNA degradation of the RNA template and improve the efficiency of reverse transcription. [000341] In some embodiments, the engineered reverse transcriptases or the recombinant RT proteins of the present disclosure can have a reduced or substantially reduced RNase H activity. The reduction or substantial reduction or complete removal of the RNase H activity of a reverse transcriptase (e.g., MMLV) can prevent the degradation of an RNA template before the initiation of the RT reaction, thereby improving the efficiency of reverse transcription. [000342] In some embodiments, the engineered reverse transcriptases or the recombinant RT proteins of the present disclosure substantially lacks RNase H activity. In that embodiment, the engineered reverse transcriptases or the recombinant RT proteins of the present disclosure can have less than 10%, 5%, 1 %, 0.5%, or 0.1 % of the RNAse H activity of a wild type enzyme or a variant having the amino acid of SEQ ID NO: 1. In some embodiments, the engineered reverse transcriptases or the recombinant RT proteins of the present disclosure lack RNase H activity. In that embodiment, the engineered reverse transcriptases or the recombinant RT proteins of the present disclosure have undetectable RNase H activity or have an RNase H activity that is less than about 1%, 0.5%, or 0.1% of the RNase H activity of a wild-type enzyme or a variant comprising the amino acid of SEQ ID NO: 1. [000343] As used herein, the term "reduced RNase H activity” means that the enzyme has less than 50%, e.g., less than 40%, less than 30%, less than 25%, or less than 20%, more preferably less than 15%, less than 10%, or less than 7.5%, and most preferably less than 5% or less than 2% of the RNase H activity of the corresponding wild type enzyme or a variant comprising the amino acid of SEQ ID NO: 1. The RNase H activity of an enzyme may be determined by assays known in the art. In some embodiments, the engineered reverse transcriptases or the recombinant RT proteins that have reduced and/or abolished RNase H activity comprise a D524 mutation in SEQ ID NO: 1 or 7. 3. Transcription efficiency [000344] In some embodiments, the engineered reverse transcriptase polypeptide or recombinant RT protein disclosed herein exhibits enhanced transcription efficiency when 93 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    compared to the transcription efficiency of a reverse transcriptase having the amino acid sequence set forth in SEQ ID NO:1 or SEQ ID NO: 7 or any RT lacking a DNA binding domain. As noted herein, the conversion of mRNA into cDNA by reverse transcriptase-mediated reverse transcription is an essential step in single cell profiling and gene expression analyses. However, the use of an unmodified reverse transcriptase to catalyze reverse transcription is inefficient for all the reasons disclosed herein. The engineered reverse transcriptases or recombinant RT proteins of the disclosure are preferably modified or mutated such that the transcription efficiency of the engineered RT polypeptide or recombinant RT protein is increased or enhanced. [000345] Further, engineered reverse transcription polypeptide or recombinant RT proteins of the present disclosure may have an unexpectedly greater ability to associate or bind to full-length transcripts (e.g., in T-cell receptor paired transcriptional profiling), as compared to that exhibited by an enzyme having the amino acid sequence set forth in SEQ ID NO:1 or non-DAT1 engineered RT. [000346] It is recognized that salt concentration, the concentration of a cell fixation chemical and/or the concentration of a process reagent in a reverse transcriptase reaction may impact function of a reverse transcriptase. For example, “chemical tolerance” is intended that an the engineered reverse transcriptase or the recombinant RT protein of the current application may exhibit a reverse transcriptase related activity in either an expanded salt concentration range or in the presence of an increased concentration of a cell fixation chemical or process reagent, or in both an expanded salt concentration range and in the presence of an increased concentration of a cell fixation chemical or process reagent, as compared to the reverse transcriptase related activity of an enzyme having the amino acid sequence set forth in SEQ ID NO: 1, 143, 145, or 172 or non-DAT1 engineered RT. [000347] An altered transcription efficiency may be an increased transcription efficiency or a decreased transcription efficiency as compared to the transcription efficiency of a reverse transcriptase having the amino acid sequence set forth in SEQ ID NO: 1. Altered transcription efficiency may be at least .1X, 0.2X, 0.3X, 0.4X, 0.5X, 0.6X, 0.7X, 0.8X, 0.9X, 1X, 1.5X, 2X, 2.5X, 3X, 3.5X, 4X, 4.5X, 5X, 5.5X, 6X, 6.5X, 7X, 7.5X, 8X, 8.5X, 9X, 10X, 15X, 20X, 25X or at least 30X greater than the transcription efficiency of a reverse transcriptase having the amino acid sequence set forth in SEQ ID NO: 1, 143, 145, or 172 or non-DAT1 engineered RT. 94 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    [000348] Transcription efficiency may be calculated as the sum of the area under the curve for the elongation, elongation plus tail, incomplete template switching (TSO) and complete template switching (TSO) regions over the total area under the curve for all products. Transcription efficiency reflects all those products for which transcription was successfully completed. Template switching oligonucleotide efficiency may be calculated as the area under the curve for the complete template switching region over the total area under the curve for all full-length products. An engineered reverse transcriptase may have an increased transcription efficiency, an increased TSO efficiency or both an increased transcription efficiency and an increased TSO efficiency. 4. Processivity [000349] In some embodiments, the engineered reverse transcriptase polypeptide or recombinant RT protein described herein possesses one or more of the following characteristics when compared to a wild-type polymerase and/or reverse transcriptase: increased thermostability; increased thermoreactivity; increased resistance to reverse transcriptase inhibitors; increased ability to reverse transcribe difficult templates; increased speed; increased processivity; increased specificity; enhanced polymerization activity; increased sensitivity, or any combination thereof. [000350] Processivity is defined as the ability of a polymerase or reverse transcriptase to carry out continuous nucleic acid synthesis on a template nucleic acid without frequent dissociation. It can be measured by the average number of nucleotides incorporated by a polymerase on a single association/disassociation event. DNA polymerase or reverse transcriptase alone produces short DNA product strand per binding event. Most DNA polymerases or reverse transcriptases are intrinsically low-processivity enzymes. The low processivity of DNA polymerase or reverse transcriptase alone is insufficient for the timely replication of a large genome. [000351] In some embodiments, the polymerization activity of the engineered reverse transcriptase polypeptide or recombinant RT protein as described herein is enhanced by about 5%, about 10%, about 15%, about 20%, about 25%, about 30%, about 35%, about 40%, about 45%, about 50%, about 55%, about 60%, about 65%, about 70%, about 75%, about 80%, about 85%, about 90%, about 90%, or about 100% as compared to the wild-type reverse transcriptase. 95 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    [000352] In some embodiments, the engineered reverse transcriptase enzyme or engineered reverse transcriptase polypeptide or recombinant RT protein reverse transcribes a RNA molecule having at least about 100, at least about 200, at least about 300, at least about 400, at least about 500, at least about 600, at least about 700, at least about 800, at least about 900, or at least about 1000 nucleotides. In another embodiment, the engineered reverse transcriptase polypeptide or recombinant RT protein reverse transcribes a RNA molecule that is at least about 1kb, at least about 2kb, at least about 3kb, at least about 4 kb, at least about 5 kb, at least about 6 kb, at least about 7 kb, at least about 8 kb, at least about 9 kb, at least about 10kb, at least about 11 kb, at least about 12 kb, at least about 13 kb, at least about 14kb, or at least about 15 kb. In another embodiment, the engineered reverse transcriptase polypeptide or recombinant RT protein reverse transcribes a RNA molecule that is at least about 7kb or at least about 8kb. [000353] In some embodiments, the increase in thermoreactivity, resistance to reverse transcriptase inhibitors, ability to reverse transcribe difficult templates, speed, processivity, specificity, or sensitivity of the engineered reverse transcriptase polypeptide or recombinant RT protein as described herein has is about 5%, about 10%, about 15%, about 20%, about 25%, about 30%, about 35%, about 40%, about 45%, about 50%, about 55%, about 60%, about 65%, about 70%, about 75%, about 80%, about 85%, about 90%, about 90%, or about 100% as compared to the wild-type polymerase. [000354] In some embodiments, the enhanced reverse transcriptase activity is an increased binding affinity and template switching efficiency as compared to a reverse transcriptase having the amino acid sequence set forth in SEQ ID NO:1, 143, 145, or 172 or non-DAT1 engineered RT. In some embodiments, the enhanced reverse transcriptase activity is an enhanced processivity as compared to the processivity of a reverse transcriptase having the amino acid sequence set forth in SEQ ID NO: 1, 143, 145, or 172 or non-DAT1 engineered RT.   [000355] Processivity relates to a reverse transcriptase’s ability to remain associated with the template while incorporating nucleotides. Measurements of processivity may include but are not limited to the number of nucleotides incorporated in a single binding event of a reverse transcriptase molecule. Processivity also relates to the affinity of the enzyme for the substrate; thus, an enzyme with increased processivity may be more resistant to the presence of an inhibitor. 96 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    IV. NUCLEIC ACIDS AND EXPRESSION VECTORS A. Nucleic Acids [000356] One aspect of the present disclosure provides an isolated nucleic acid molecule encoding the engineered reverse transcriptase, the recombinant RT protein or derivatives thereof as described herein. One aspect of the present disclosure provides an isolated nucleic acid molecule encoding any of the engineered RT polypeptides or the recombinant RT protein described herein. [000357] In some embodiments, the engineered reverse transcriptase polypeptide or the recombinant RT protein disclosed herein can been coded by a nucleic acid set forth herein or readily derived in light of polypeptide information provided herein and known in the art. In some embodiments, the isolated nucleic acid molecule encoding the RT polypeptide can comprise a sequence selected from SEQ ID NO; 25; SEQ ID NO: 136, SEQ ID NO:138, SEQ ID NO:140, SEQ ID NO:142, SEQ ID NO:144, SEQ ID NO:146, SEQ ID NO:148, SEQ ID NO:150, SEQ ID NO:152, SEQ ID NO:154, SEQ ID NO:156, SEQ ID NO:158, SEQ ID NO:160, SEQ ID NO:162, SEQ ID NO:164, SEQ ID NO:167, SEQ ID NO:169, or SEQ ID NO: 171; or a non- limiting embodiment of a nucleic acid sequence of Table 1 or Table 2. [000358] The reverse transcriptase polypeptides, or the DNA binding domains need not be encoded by any specific nucleic acid exemplified herein. For example, redundancy in the genetic code allows for variations in nucleotide codon sequences that nevertheless encode the same amino acid. Accordingly, engineered polymerases of the present disclosure can be produced from nucleic acid sequences that are different from those set forth herein, for example, being codon optimized for a particular expression system. Codon optimization can be carried out, for example, as set forth in Athey et al., BMC Bioinformatics, 18:391-401 (2017). [000359] Wild type nucleic acids (e.g., RT or DNA binding domain) may be isolated from naturally occurring sources to be used as starting material to generate novel polymerases. Generally, the nomenclature and the laboratory procedures in recombinant DNA technology described below are those well-known and commonly employed in the art. Standard techniques for cloning, DNA and RNA isolation, amplification and purification are known. Generally enzymatic reactions involving DNA ligase, DNA polymerase, restriction endonucleases are the 97 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    like are performed according to the manufacturer's specifications. These techniques and various other techniques are generally performed according to Sambrook & Russell, Molecular Cloning- A Laboratory Manual, Cold Spring Harbor Laboratory, Cold Spring Harbor, N.Y., (1989) or Ausubel et al., Current Protocols in Molecular Biology, Vol.1-3, John Wiley & Sons, Inc. (1994-1998). [000360] The isolation of nucleic acids (e.g., RT or DNA binding domain) may be accomplished by a variety of techniques. The nucleic acids of the present disclosure can be generated from the wild type sequences. The wild type sequences are altered to create modified sequences. Wild type molecules (e.g., RT or DNA binding domain) can be modified to create the engineered RT described in the present application using methods that are well known in the art. Exemplary modification methods are site-directed mutagenesis, point mismatch repair, or oligonucleotide-directed mutagenesis. B. Vectors [000361] Another aspect of the present disclosure provides an expression vector comprising the isolated nucleic acid encoding the engineered reverse transcriptase polypeptides or derivatives thereof as described herein. A “vector” refers to a polynucleotide, which when independent of the host chromosome, is capable replication in a host organism. Preferred vectors include plasmids and typically have an origin of replication. Vectors can comprise, e.g., transcription and translation terminators, transcription and translation initiation sequences, and promoters useful for regulation of the expression of the particular nucleic acid. The polymerases of the present disclosure can be expressed in a variety of host cells, including E. coli, other bacterial hosts, yeasts, filamentous fungi, and various higher eukaryotic cells such as the COS, CHO and HeLa cells lines and myeloma cell lines. Techniques for gene expression in microorganisms are described in, for example, Smith, Gene Expression in Recombinant Microorganisms (Bioprocess Technology, Vol.22), Marcel Dekker, 1994. Examples of bacteria that are useful for expression include, but are not limited to, Escherichia, Enterobacter, Azotobacter, Erwinia, Bacillus, Pseudomonas, Klebsielia, Proteus, Salmonella, Serratia, Shigella, Rhizobia, Vitreoscilla, and Paracoccus. Filamentous fungi that are useful as expression hosts include, for example, the following genera: Aspergillus, Trichoderma, Neurospora, Penicillium, Cephalosporium, Achlya, Podospora, Mucor, Cochliobolus, and Pyricularia. Synthesis of heterologous proteins in yeast is 98 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    well known and described in the literature. There are many expression systems for producing the polymerase polypeptides of the present invention that are well known to those of ordinary skill in the art. C. Host cells [000362] Another aspect of the present disclosure provides a host cell transfected with the expression vector comprising the isolated nucleic acid encoding the engineered reverse transcriptase as described herein. Eukaryotic expression systems for mammalian cells, yeast, and insect cells are well known in the art and are also commercially available. In yeast, vectors include Yeast Integrating plasmids (e.g., YIp5) and Yeast Replicating plasmids (the YRp series plasmids) and pGPD-2. Expression vectors containing regulatory elements from eukaryotic viruses are typically used in eukaryotic expression vectors, e.g., SV40 vectors, papilloma virus vectors, and vectors derived from Epstein-Barr virus. Other exemplary eukaryotic vectors include pMSG, pAV009/A+, pMTO10/A+, pMAMneo-5, baculovirus pDSVE, and any other vector allowing expression of proteins under the direction of the CMV promoter, SV40 early promoter, SV40 later promoter, metallothionein promoter, murine mammary tumor virus promoter, Rous sarcoma virus promoter, polyhedrin promoter, or other promoters shown effective for expression in eukaryotic cells. [000363] Once expressed, the engineered reverse transcriptase or a derivative thereof can be purified according to standard procedures of the art, including ammonium sulfate precipitation, affinity purification columns, column chromatography, gel electrophoresis and the like. Substantially pure compositions of at least about 90 to about 95% homogeneity are preferred, and about 98 to about 99% or more homogeneity are most preferred. Once purified, partially or to homogeneity as desired, the polypeptides may then be used (e.g., as immunogens for antibody production). [000364] To facilitate purification of the engineered reverse transcriptase or a derivative thereof, the nucleic acids that encode the engineered reverse transcriptase or derivatives thereof can also include a coding sequence for an epitope or “tag” for which an affinity binding reagent is available. Examples of suitable epitopes include the myc and V-5 reporter genes; expression vectors useful for recombinant production of fusion polypeptides having these epitopes are commercially available (e.g., Invitrogen (Carlsbad Calif.) vectors pcDNA3.1/Myc-His and 99 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    pcDNA3.1/V5-His are suitable for expression in mammalian cells). Additional expression vectors suitable for attaching a tag to the fusion proteins of the disclosure, and corresponding detection systems are known to those of skill in the art as described herein, and several are commercially available (e.g., FLAG″ (Kodak, Rochester N.Y.). Another example of a suitable tag is a polyhistidine sequence, which is capable of binding to metal chelate affinity ligands. Typically, six adjacent histidines are used (6His-tag, his-tag), although one can use more or less than six. Suitable metal chelate affinity ligands that can serve as the binding moiety for a polyhistidine tag include nitrilo-tri-acetic acid (NTA). [000365] One of skill in the art would recognize that after biological expression or purification, the engineered reverse transcriptase or derivatives thereof may possess a conformation substantially different than the native conformations of the constituent polypeptides. In this case, it may be necessary or desirable to denature and reduce the engineered reverse transcriptase or a derivative thereof and cause the engineered reverse transcriptase or a derivative thereof to re-fold into the preferred conformation. Methods of reducing and denaturing proteins and inducing re- folding are well known to those of skill in the art. V. COMPOSITIONS AND REACTION MIXTURES [000366] The present disclosure further provides compositions comprising a variety of components in various combinations needed for nucleic acid amplification using the engineered RT polypeptides or recombinant proteins disclosed herein. [000367] One aspect of the present disclosure provides a composition comprising any of the recombinant RT proteins described herein. One aspect of the present disclosure provides a composition comprising any of the engineered RT polypeptides described herein. One aspect of the present disclosure provides a composition comprising any of the engineered RT polypeptides described herein. One aspect of the present disclosure provides a composition comprising any of the expression vectors described herein. One aspect of the present disclosure provides a composition comprising the host cells described herein. In some embodiments, any one of the compositions described herein further comprise a buffer. [000368] In some embodiments of the present disclosure, the compositions are formulated by admixing one or more engineered reverse transcriptase polypeptides or recombinant RT proteins, 100 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    or derivatives thereof of the present disclosure in a buffered salt solution. One or more DNA polymerases and/or one or more nucleotides, and/or one or more primers may optionally be added to create the compositions of the invention. These compositions can be used in the methods disclosed herein to produce, analyze, quantitate and otherwise manipulate nucleic acid molecules (e.g., using reverse transcription or one-step RT-PCR procedures). [000369] In some embodiments, the engineered reverse transcriptases or the recombinant RT proteins disclosed herein are provided at working concentrations (e.g., 1×) in stable buffered salt solutions. [000370] The terms “stable” and “stability” as used herein generally mean the retention by a composition, such as an enzyme (e.g., engineered reverse transcriptase or the recombinant RT protein) composition, of at least 70%, preferably at least 80%, and most preferably at least 90%, of the original enzymatic activity (in units) after the enzyme (e.g., engineered reverse transcriptase or the recombinant RT protein) or composition containing the enzyme has been stored for about one week at a temperature of about 4° C, about two to six months at a temperature of about −20° C, and about six months or longer at a temperature of about −80° C. [000371] As used herein, the term “working concentration” means the concentration of an enzyme (e.g., engineered reverse transcriptase or the recombinant RT protein) that is at or near the optimal concentration used in a solution to perform a particular function such as reverse transcription of nucleic acids. [000372] Such compositions can also be formulated as concentrated stock solutions (e.g., 2×, 3×, 4×, 5×, 6×, 10×, etc.). In some embodiments, having the composition as a concentrated (e.g., 5x) stock solution allows a greater amount of nucleic acid sample to be added (such as, for example, when the compositions are used for nucleic acid synthesis). The water used in forming the compositions of the present invention is preferably distilled, deionized and sterile filtered (through a 0.1-0.2 micrometer filter) and is free of contamination by DNase and RNase enzymes. Such water is available commercially, for example from Life Technologies (Carlsbad, Calif.) or may be made as needed according to methods well known to those skilled in the art. 101 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    VI. METHODS FOR USING ENGINEERED REVERSE TRANSCRIPTASES [000373] The engineered reverse transcriptases of the present disclosure may be used in any application in which a reverse transcriptase with the indicated altered activity is desired. Methods of using reverse transcriptases are known in the art and one skilled in the art may select any of the engineered reverse transcriptases disclosed herein. In some embodiments, a reverse transcription reaction introduces a barcode. In some embodiments, the barcoding reaction is an enzymatic reaction. In some embodiments, the barcoding reaction is a reverse transcription amplification reaction that generates complementary deoxyribonucleic acid (cDNA) molecules upon reverse transcription of ribonucleic acid (RNA) molecules of the cell. In some embodiments, the RNA molecules are released from the cell. In some embodiments, the RNA molecules are released from the cell by lysing the cell. In some embodiments, the RNA molecules are messenger RNA (mRNA). A. Amplification Methods [000374] One aspect of the present disclosure provides a method for performing a reverse transcription reaction for generating a nucleic acid product from an RNA template using an engineered reverse transcriptase or recombinant RT protein described herein. The engineered reverse transcriptases or recombinant RT protein of the present application may be used in any application in which a reverse transcriptase with the indicated altered activity is desired. Methods of using reverse transcriptases are known in the art. One skilled in the art may select any of the engineered reverse transcriptases disclosed herein. [000375] One aspect of the present disclosure provides a method for performing a reverse transcription reaction for generating a nucleic acid product from an RNA template comprising contacting under suitable conditions a biological sample or extract thereof with an engineered RT polypeptide, or a recombinant RT protein described herein. The cell can be fixed. The cell can be permeabilized. Alternatively, the cell can be permeabilized and fixed. In some embodiments, the cell is a cell bead. In some embodiments, the cell bead is fixed. In some embodiments, the nucleus is permeabilized or fixed. In some embodiments, the nucleus is permeabilized and fixed. In some embodiments, the biological sample comprises a suitable cellular preparation selected from cell populations and/or single cells. In some embodiments, the biological sample comprises a tissue. 102 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    [000376] In some embodiments, when the biological sample is a cell, a cell bead, or a nucleus, the reverse transcription reaction is part of a single cell RNA sequencing assay. In one embodiment, the single cell RNA sequencing assay further comprises, prior to the reverse transcription, partitioning the cell, the cell bead, or the nucleus into a partition. In one embodiment, the single cell RNA sequencing assay further comprises, after the reverse transcription reaction, hybridizing the nucleic acid product to an oligonucleotide molecule comprising a partition-specific barcode. [000377] In some embodiments, when the biological sample is a cell or tissue sample immobilized on a surface, the reverse transcription reaction is part of a spatial RNA sequencing assay. [000378] In some embodiments of the method described herein, the engineered RT polypeptide or the recombinant RT protein enhances template switching (TS) efficiency, processivity efficiency, binding affinity, transcription efficiency, chemical tolerance, ability to yield mitochondrial unique molecular identity (UMI) counts, ability to yield ribosomal unique molecular identity (UMI) counts, strand displacement, end-to-end template jumping, or any combination thereof, when compared to an RT polypeptide or a recombinant RT protein lacking a conjugated DNA binding domain. [000379] In some embodiments, the engineered RT polypeptide or the recombinant RT protein enhances at least two or more of template switching (TS) efficiency, processivity efficiency, binding affinity, transcription efficiency, chemical tolerance, ability to yield mitochondrial unique molecular identity (UMI) counts, strand displacement, end-to-end template jumping, or ability to yield ribosomal unique molecular identity (UMI) counts, when compared to an RT polypeptide or a recombinant RT protein lacking a conjugated DNA binding domain. [000380] In some embodiments, the recombinant RT protein or the engineered RT enhances transcript capture during amplification, when compared to an RT polypeptide or a recombinant RT protein lacking a conjugated DNA binding domain. [000381] In some embodiments, the DNA binding domain of the recombinant RT protein or the engineered RT enhances the hybridization of a transcript and a primer during the amplification 103 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    process, when compared to an RT polypeptide or a recombinant RT protein lacking a conjugated DNA binding domain. [000382] The engineered RT polypeptide or the recombinant RT protein can comprise a DNA binding domain comprising an amino acid sequence selected from SEQ ID NO:2, 3, 5, 6, 8, 9, or 11-24; and an amino acid sequence selected from SEQ ID NOs: 27-61, 63, 65, 67, 69, 71, 73, 75, 77, 79, 81, 83, 85, 87, 89, 91, 93, 95, 97, 99, 101, 103, 105, 107, 109, 111, 113, 115, 117, 119, 121, 123, 125, 127, 129, 131, 133, 141, 143, 145, 147, 149, 151, 157, 159, 172, or 173. [000383] In some embodiments, the amino acid sequence of the engineered RT polypeptide or the recombinant RT protein comprises an amino acid sequence having at least about 90%, at least about 92%, at least about 95%, at least about 97%, at least about 98%, or at least about 99% sequence identity to SEQ ID NO: 174-188. [000384] The engineered RT polypeptide or the recombinant RT protein can comprise M39V, M66I, Q91R, I347V, and H594Q substitution in SEQ ID NO: 143. The engineered RT polypeptide or the recombinant RT protein can comprise SEQ ID NO: 129 (SOLD 034). The engineered RT polypeptide or the recombinant RT protein can comprise SOLD 001 (SEQ ID NO: 65). The engineered RT polypeptide or the recombinant RT protein can comprise SOLD 33 VDG (SEQ ID NO: 173). In some embodiments, the engineered RT polypeptide or the recombinant RT protein can comprise M39V, T542D, D583N, E607G, A644V, D653H, K658R, and L671P in SEQ ID NO: 1 or 143. [000385] The engineered RT polypeptide or the recombinant RT protein can comprise M39V, T542D, D583N, E607G, A644V, D653H, K658R, and/or L671P in SEQ ID NO: 143. The engineered RT polypeptide or the recombinant RT protein can comprise SEQ ID NO: 111 (SOLD 025). The engineered RT or the recombinant RT protein can comprise a M39V, M66I, Q91R, I347V, and/or H594Q in SEQ ID NO: 143. [000386] In some embodiments, the engineered RT polypeptide or the recombinant RT protein can comprise an amino acid sequence that is at least about 90%, at least about 92%, at least about 95%, at least about 97%, at least about 98%, or at least about 99% identical to an amino acid sequence disclosed in Table 1 or Table 2. 104 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    [000387] One aspect of the present disclosure provides a method of using the engineered RT polypeptide, or the recombinant RT protein described herein, the method comprising contacting the engineered RT polypeptide or the recombinant RT protein with a nucleic acid template under suitable conditions to produce a polymerized nucleic acid product. In some embodiments, the nucleic acid template comprises an RNA, a DNA, or a nucleic acid comprising an unnatural nucleotide. In some embodiments, the nucleic acid template comprises an RNA. [000388] In some embodiments, the engineered reverse transcriptase comprises an M39 mutation, a K47 mutation, an L435 mutation, a D449 mutation, a D524 mutation, an E607 mutation, a D653 mutation and an L671 mutation in SEQ ID NO:7. In some embodiments, the engineered reverse transcriptase comprises a mutation selected from a K13 mutation, a K13L mutation, a D36 mutation, an N37 mutation, a V2 mutation, a D36L mutation, an insertion, and a combination thereof. [000389] In some embodiments, the engineered reverse transcriptases or the recombinant RT proteins, or derivatives thereof of the present disclosure are used in reverse transcription reactions, such as RT-PCR, or other known reactions in the art where nucleic acids, for example RNA molecules, are reverse transcribed using a reverse transcriptase. [000390] The engineered reverse transcriptase, the recombinant RT protein or a derivative thereof as described herein may be used to make nucleic acid molecules from one or more templates. Such methods can comprise mixing one or more nucleic acid templates (e.g., RNA, such as non-coding RNA (ncRNA), messenger RNA (mRNA), micro RNA (miRNA), and small interfering RNA (siRNA) molecules) with one or more of the engineered reverse transcriptases of the disclosure and incubating the mixture under conditions sufficient to generate one or more nucleic acid molecules complementary to all or a portion of the one or more nucleic acid templates. Other methods of cDNA synthesis which may advantageously use the engineered reverse transcriptase or the recombinant RT protein of the present disclosure will be readily apparent to one of ordinary skill in the art. [000391] In some embodiments, the method of using the engineered reverse transcriptase, or the recombinant RT protein or a derivative thereof as described herein comprises the amplification of one or more nucleic acid molecules comprising mixing one or more nucleic acid templates with one of the engineered reverse transcriptase polypeptide or recombinant RT 105 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    proteins or a derivative thereof of the disclosure, and incubating the mixture under conditions sufficient to amplify one or more nucleic acid molecules complementary to all or a portion of the one or more nucleic acid templates. In one embodiment, the method may comprise the use of one or more DNA polymerases and may be employed as in standard reverse transcription-polymerase chain reaction (RT-PCR) reactions. [000392] In some embodiments, the method of using the engineered reverse transcriptase, recombinant RT protein or a derivative thereof as described herein may be one-step (e.g., one- step RT-PCR) or two-step (e.g., two-step RT-PCR) reactions. In one embodiment, the one-step RT-PCR type reactions may be accomplished in one tube thereby lowering the possibility of contamination. Such one-step reactions can comprise (a) mixing a nucleic acid template (e.g., mRNA) with one or more engineered reverse transcriptase polypeptides or recombinant RT proteins or derivatives thereof of the present disclosure and one or more polymerases and (b) incubating the mixture under conditions sufficient to amplify a nucleic acid molecule complementary to all or a portion of the template. [000393] In another embodiment, two-step RT-PCR reactions may be accomplished in two separate steps. Such methods can comprise (a) mixing a nucleic acid template (e.g., mRNA) with an engineered reverse transcriptase polypeptide or a recombinant RT protein or a derivative thereof of the present disclosure, (b) incubating the mixture under conditions sufficient to make a nucleic acid molecule (e.g., a DNA molecule) complementary to all or a portion of the template, (c) mixing the nucleic acid molecule with one or more DNA polymerases and (d) incubating the mixture of step (c) under conditions sufficient to amplify the nucleic acid molecule. For amplification of long nucleic acid molecules (i.e., greater than about 3-5 kb in length), a combination of DNA polymerases and the engineered reverse transcriptase polypeptide or recombinant RT protein or a derivative thereof of the present disclosure may be used. [000394] Amplification methods which may be used in accordance with the present invention (e.g., using one or more engineered reverse transcriptase polypeptides or recombinant RT proteins or derivatives thereof of the present disclosure) include PCR, Isothermal Amplification, Strand Displacement Amplification (SDA), and Nucleic Acid Sequence-Based Amplification (NASBA); as well as more complex PCR-based nucleic acid fingerprinting techniques such as Random Amplified Polymorphic DNA (RAPD) analysis, Arbitrarily Primed PCR (AP-PCR) 106 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    DNA Amplification Fingerprinting (DAF); microsatellite PCR; Directed Amplification of Minisatellite-region DNA (DAVID); digital droplet PCT (ddPCR) and Amplification Fragment Length Polymorphism (AFLP) analysis. In some embodiments, the engineered reverse transcriptase disclosed herein may be used in methods of amplifying or sequencing a nucleic acid molecule comprising one or more polymerase chain reactions (PCRs), such as any of the PCR- based methods described above. [000395] Methods of producing an engineered reverse transcriptase, an engineered reverse transcriptase or a derivative thereof of the present disclosure are known to those of skill in the art of molecular biology or molecular genetics. For example, nucleic acids encoding the wild-type polymerase or nucleic acid binding domains can be generated using routine techniques in the field of recombinant genetics. B. Nucleic Acid Sample Processing [000396] Another aspect of the present disclosure provides a nucleic acid extension method comprising contacting a target nucleic acid molecule with an engineered reverse transcriptase or a recombinant RT protein and a plurality of nucleic acid barcoded molecules comprising a barcode sequence, and incubating the target nucleic acid, the engineered reverse transcriptase or the recombinant RT protein and barcoded molecules under conditions in which the barcoded molecules are extended by the engineered reverse transcriptase or the recombinant RT protein. In some embodiments, the engineered reverse transcriptase or the recombinant RT protein comprises the amino acid sequence of an engineered RT or an recombinant RT protein described herein or a derivatives thereof. The target nucleic acid hybridizes to one of the plurality of barcoded molecules and the hybridized barcoded molecule is extended by the engineered reverse transcriptase or the recombinant RT protein described herein. [000397] The novel engineered reverse transcriptase polypeptide or the recombinant RT protein described herein can be used to generate a Single Cell 3' (SC-3') and/or 5’ (SC-5') gene expression libraries. The SC-3' and SC-5' assays are similar but capture different ends of the polyadenylated transcript in the final library. Both solutions use poly-dT primer for reverse transcription). In the SC-3' assay (FIGs 14B), the poly-dT sequence is located on the gel bead oligo. In the SC-5' assay (FIGs 14A), the poly-dT is supplied as an RT primer. A template switching oligo (TSO) is used in both assays to reverse transcribe the full-length transcript. 107 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    [000398] After amplifying the cDNA, transcripts are randomly fragmented under conditions that favor 300-400 bp length fragments. Downstream of fragmentation, only transcripts containing both (1) a 10x Barcode and (2) an Illumina® Read 2 adaptor, which is ligated on to the cDNA after fragmentation, can be amplified during the Sample Index PCR. This results in final 10x libraries that either represent the 3' end of the transcript (as the 10x Barcode is adjacent to the polyA tail on the 3' end of the transcript) or the 5' end of the transcript (as the 10x Barcode is adjacent to the TSO and the 5' end of the transcript). See e.g., kb.10xgenomics.com/hc/en-us/articles/360000939852-What-is-the-difference-between-Single- Cell-3-and-5-Gene-Expression-libraries-. 1. RNA Template [000399] In some embodiments, the nucleic acid is a ribonucleic acid (RNA) molecule; and the engineered reverse transcriptase polypeptide or recombinant RT protein reverse transcribes the RNA molecule thereby generating a first strand cDNA. [000400] A first strand cDNA reaction can be optionally performed using template switching oligonucleotides. For example, a template switching oligonucleotide can hybridize to a poly(C) tail added to a 3’ end of the cDNA by the engineered reverse transcriptase polypeptide or recombinant RT protein described herein. The original mRNA template and template switching oligonucleotide can then be denatured from the cDNA and a barcoded capture probe can then hybridize with the cDNA and a complement of the cDNA can be generated. The first strand cDNA can then be purified and collected for downstream amplification steps. The first strand cDNA can be amplified using PCR, where the forward and reverse primers flank the spatial barcode and target analyte regions of interest, generating a library associated with a particular spatial barcode. In some embodiments, the cDNA comprises a sequencing by synthesis (SBS) primer sequence. The library amplicons are sequenced and analyzed to decode spatial information. [000401] Exemplary steps for sample preparation, permeabilization, DNA generation (e.g., first strand cDNA generation and second strand generation), DNA amplification (e.g., cDNA amplification) and quality control, and spatial gene expression library construction are disclosed for example in WO 2020/047002, WO 2020/047004, WO 2020/047005, WO 2020/047007, and WO 2020/047010, all of which are incorporated herein by reference in their entireties. 108 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    [000402] In some embodiments, a reverse transcription reaction introduces a barcode. In some embodiments, the barcoding reaction is an enzymatic reaction. In some embodiments, the barcoding reaction is a reverse transcription amplification reaction that generates complementary deoxyribonucleic acid (cDNA) molecules upon reverse transcription of ribonucleic acid (RNA) molecules of the cell. In some embodiments, the RNA molecules are released from the cell. In some embodiments, the RNA molecules are released from the cell by lysing the cell. In some embodiments, the RNA molecules are released from the cell by permeabilizing the cell, or a tissue which comprises a plurality of the same and/or different cell types. In some embodiments, the RNA molecules are messenger RNA (mRNA). [000403] In some embodiments, a reverse transcription reaction using the engineered reverse transcriptase, the recombinant RT protein or derivative thereof of the present disclosure is initiated at the point of hybridization of the capture sequences to the RNA molecules, with the capture probe being extended by the engineered reverse transcriptase polypeptide or recombinant RT protein of the present disclosure in a template directed fashion using the hybridized mRNA as a template. The recombinant RT protein or the engineered RT polypeptide can exhibit increased transcript capture during amplification. In that embodiment, the DNA binding domain of the engineered Rt polypeptide or recombinant RT protein can enhance the hybridization of a transcript and a primer during a nucleic acid amplification process. In some embodiments, the primer can comprise a poly-dT or a poly(dT)VN sequence and a non-poly(dT) sequence; and the transcript can comprise a poly-dA sequence. The DNA binding domain (e.g., DAT1 or variant thereof) can stabilize the oligo(A)-olgo(T) based transcript-primer complex during a nucleic acid amplification process. In some embodiments, the primer is a barcoded molecule. [000404] In some embodiments, the reverse transcription reaction produces single stranded cDNA molecules each having a molecular tag and barcode associated with the cDNA, followed by amplification of cDNA to produce a double stranded cDNA that includes the sequences of the barcoded molecules. [000405] In some embodiments, the plurality of nucleic acid barcoded molecules comprise an oligo(dT) sequence. In that embodiment, the engineered reverse transcriptase polypeptide or recombinant RT protein reverse transcribes the mRNA molecule into a complementary DNA molecule using the mRNA hybridized to the oligo(dT) sequence of the nucleic acid barcoded 109 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    molecules as a template, and the nucleic acid binding domain binds and stabilizes the mRNA- oligo(dT) hybrid during the reverse transcription. Following reverse transcription, the engineered reverse transcriptase polypeptide or recombinant RT protein as described herein further amplifies the complementary DNA molecule comprising the barcode sequence, thereby generating an amplified DNA product comprising the barcode sequence, molecular tag sequence, or complements thereof. [000406] In some embodiments of the nucleic acid extension method described herein, the method can comprise a second nucleic acid molecule comprising an oligo(dT) sequence. In that embodiment, the plurality of nucleic acid barcoded molecules comprise an oligo(dT) sequence; and the nucleic acid binding domain of the engineered reverse transcriptase polypeptide or recombinant RT protein binds and stabilizes the mRNA-Oligo(dT) hybrid, while the polymerase domain of the engineered reverse transcriptase polypeptide or recombinant RT protein reverse transcribes the mRNA molecule using the second nucleic acid molecule comprising the oligo(dT) sequence, thereby generating a complementary DNA molecule. In this embodiment, the engineered reverse transcriptase polypeptide or recombinant RT protein further amplifies the complementary DNA molecule, thereby generating an amplified DNA product comprising a barcode sequence. [000407] In some embodiments, the nucleic acid extension method comprises a cell, a population of cells, or a tissue and the template nucleic acid molecule is from the cell, population of cells or the tissue. [000408] In some embodiments, the molecular tags are coupled to priming sequences and the barcoding reaction is initiated by hybridization of the priming sequences to the RNA molecules. In some embodiments, each priming sequence comprises a random N-mer sequence. In some embodiments, the random N-mer sequence is complementary to a 3’ sequence of a ribonucleic acid molecule of the cell. In some embodiments, the random N-mer sequence comprises a poly- dT sequence having a length of at least 5 bases. In some embodiments, the random N-mer sequence comprises a poly-dT sequence having a length of at least 10 bases. [000409] In some embodiments, the barcoding reaction is performed by extending the priming sequences in a template directed fashion using reagents for reverse transcription. In some embodiments, the reagents for reverse transcription comprise a reverse transcription enzyme 110 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    (e.g., engineered RT polypeptide or recombinant RT protein), a buffer and a mixture of nucleotides. In some embodiments, the reverse transcription enzyme adds a plurality of non- template oligonucleotides upon reverse transcription of a ribonucleic acid molecule. In some embodiments, the reverse transcription enzyme is an engineered RT polypeptide or recombinant RT protein as disclosed herein. [000410] In some embodiments, the barcoding reaction produces single stranded complementary deoxyribonucleic acid (cDNA) molecules each having a molecular tag from said molecular tags on a 5’ end thereof, followed by amplification of cDNA to produce a double stranded cDNA having the molecular tag on the 5’ end and a 3’ end of the double stranded cDNA. [000411] In some embodiments, a molecular tag which comprises a barcode plus additional functional sequences, or only additional functional sequences, is further included into a cDNA molecule generated during a reverse transcription reaction. In some embodiments, the reagents for reverse transcription comprise a reverse transcription enzyme (e.g., the engineered reverse transcriptase or the recombinant RT protein described herein), a buffer, and a mixture of nucleotides. In some embodiments, the reverse transcription enzyme adds a plurality of non- template oligonucleotides upon reverse transcription of a ribonucleic acid molecule from the nucleic acid molecules. In some embodiments, the reverse transcription enzyme is an engineered reverse transcriptase or a recombinant RT protein as disclosed herein. [000412] In one aspect, the present disclosure provides methods that utilize the engineered reverse transcriptase polypeptides or the recombinant RT protein described herein for nucleic acid sample processing. In one embodiment, the method comprises contacting a template ribonucleic acid (RNA) molecule with an engineered reverse transcriptase to reverse transcribe the RNA molecule to a complementary DNA (cDNA) molecule. The contacting step may be in the presence of a plurality of nucleic acid barcode molecules, wherein each nucleic acid barcode molecule comprises a barcode sequence. The nucleic acid barcode molecule may comprise a sequence configured to couple to a template RNA molecule. Suitable sequences include, without limitation, an oligo(dT) sequence, a random N-mer primer, or a target-specific primer. The nucleic acid barcode molecule may comprise a template switching sequence. 111 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    [000413] In other embodiments, the RNA molecule is a messenger RNA (mRNA) molecule. In one embodiment, the contacting step provides conditions suitable to allow the engineered reverse transcriptase to: (i) transcribe the mRNA molecule into the cDNA molecule with the oligo(dT) sequence and/or (ii) perform a template switching reaction, thereby generating the cDNA molecule which comprises the barcode sequence, or a derivative thereof. [000414] In another embodiment, the contacting step may occur in (i) a partition having a reaction volume (e.g., as further described herein and see e.g., US Patent Nos.10,400,280 and 10,323,278, each of which is incorporated herein by reference in its entirety); (ii) in a bulk reaction where the reaction components (e.g., template RNA and engineered reverse transcriptase) are in solution; or (iii) on a nucleic acid array (see e.g., US Patent Nos.10,480,022 and 10,030,261 as well as WO/2020/047005 and WO/2020/047010, each of which is incorporated herein by reference in its entirety). Further, the reverse transcription reaction may occur in a tissue (e.g., in situ reverse transcription), on a template that is associated with a sequence on a substrate, such as practiced in spatial transcriptomics, or further in a RT-PCR or other reverse transcription reaction in vitro on a purified target, partially purified target or unpurified target as found for example in a cellular lysate. [000415] Examples of assays involving nucleic acid sample processing may include, but are not limited to, single-cell transcription profiling, single-cell sequence analysis, immune profiling of individual T and B cells, single-cell chromatin accessibility analysis (e.g., ATAC seq analysis), single cell processing and analysis, paired single cell TCR sequencing, paired TCRα and TCRβ. These exemplary assays may be carried out using commercially available systems for encapsulating biological samples, gel beads, barcodes, and/or other compounds/materials in droplets, such as The Chromium System (10X Genomics, Pleasanton CA USA). Engineered RT polypeptide or recombinant RT protein may be used in methods of profiling a T-Cell receptor (TCR). [000416] In various embodiments, the poly-dT sequence may be extended in a reverse transcription reaction using the mRNA as a template to produce a cDNA transcript complementary to the mRNA and also includes sequence of a barcode oligonucleotide. Terminal transferase activity of the reverse transcriptase can add additional bases to the cDNA transcript (e.g., polyC). The switch oligo may then hybridize with the additional bases added to the cDNA 112 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    transcript and facilitate template switching. A sequence complementary to the switch oligo sequence can then be incorporated into the cDNA transcript via extension of the cDNA transcript using the switch oligo as a template. Within any given partition, all the cDNA transcripts of the individual mRNA molecules include a common barcode sequence. However, by including the unique random N-mer sequence, the transcripts made from different mRNA molecules within a given partition will vary at this unique sequence. As described elsewhere herein, this provides a quantification feature that can be identifiable even following any subsequent amplification of the contents of a given partition, e.g., the number of unique segments associated with a common barcode can be indicative of the quantity of mRNA originating from a single partition, and thus, a single cell. The cDNA transcript may then be amplified with PCR primers. The amplified product may then be purified (e.g., via solid phase reversible immobilization (SPRI)). The amplified product can be ligated to additional functional sequences, and further amplified (e.g., via PCR). The functional sequences may include a sequencer specific flow cell attachment sequence such as but not limited to., a P7 sequence for Illumina® sequencing systems, as well as functional sequence, which may include a sequencing primer binding site, e.g., for a R2 primer for Illumina® sequencing systems, as well as functional sequence, which may include a sample index, e.g., an i7 sample index sequence for Illumina® sequencing systems. [000417] Although described in terms of specific sequence references used for certain sequencing systems, e.g., Illumina® systems, it will be understood that the reference to these sequences is for illustration purposes only, and the methods described herein may be configured for use with other sequencing systems incorporating specific priming, attachment, index, or other operational sequences used in those systems, e.g., systems available from Ion Torrent, Oxford Nanopore, Genia, Pacific Biosciences, Complete Genomics, and the like. 2. Volume [000418] As described herein, wild-type and variants MMLV RT are not optimal for reverse transcription of mRNA when using high throughput amplification reaction assays (e.g., spatial array and single cell transcriptomics assay) and the like. This is because high throughput amplification reaction assays require reaction volumes that are usually less than about 1 nanoliter. Accordingly, the present disclosure provides novel engineered reverse transcriptase 113 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    polypeptide or recombinant RT proteins that function efficiently in high throughput amplification reaction assays that require reaction volumes of less than about 1 nanoliter. [000419] In some embodiments, the method comprises providing a reaction volume which comprises an engineered reverse transcriptase and a template ribonucleic acid (RNA) molecule. In one other embodiment, the contacting occurs in a reaction volume, which may be less than 1 nanoliter, less than 750 picoliters, or less than 500 picoliters. In other embodiments, the reaction volume is present in a partition, such as a droplet or well (including a microwell or a nanowell). [000420] In some embodiments, the engineered reverse transcriptases, the recombinant RT protein, or derivatives thereof as described herein are used in a reaction volume less than about 1 nanoliter (nL). In some embodiments, the engineered reverse transcriptases, the recombinant RT proteins, or derivatives thereof, as described herein are used in a reaction volume that is less than about 500 picoliter (pL). In some embodiments, the reaction volume is contained within a partition. In some embodiments, the reaction volume is contained within a droplet. In some embodiments, the reaction volume is contained within a droplet in an emulsion. In some embodiments, the reaction volume is contained within a droplet emulsion having a reaction volume of less than about 1 nL. In some embodiments, the reaction volume is contained within a droplet emulsion having a reaction volume of less than about 500 pL. [000421] In some embodiments, the reaction volume is contained within a well. In some embodiments, the reaction volume is contained within a well having a reaction volume less than about 1 nL. In some embodiments, the reaction volume is contained within a well. In some embodiments, the reaction volume is contained within a well having a reaction volume less than about 500 pL. In some embodiments, the reaction volume is contained within a well in an array of wells having an extracted nucleic acid molecule, and the template nucleic acid molecule is the extracted nucleic acid molecule. In some embodiments, the reaction volume is contained within a well in an array of wells having a cell comprising a template nucleic acid molecule, and where the template nucleic acid molecule is released from the cell. [000422] In another embodiment, a method comprises providing a reaction volume, which comprises an engineered reverse transcriptase and a template ribonucleic acid (RNA) molecule and is considered a “low volume reaction”. The reaction volume may comprise a plurality of nucleic acid barcode molecules, and each nucleic acid barcode molecule comprises a barcode 114 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    sequence. In an embodiment, the contacting occurs in a reaction volume, a low volume reaction, which may be less than 1 nanoliter, less than 750 picoliters, or less than 500 picoliters. In other embodiments, the reaction volume is present in a partition, such as a droplet or well (including a microwell or a nanowell). 3. Unique molecular identifier (UMI) [000423] In some embodiments, the barcoding reaction produces single stranded complementary deoxyribonucleic acid (cDNA) molecules each having a molecular tag on a 5’ end thereof, followed by amplification of the cDNA to produce a double stranded DNA having the molecular tag on the 5’ end and a 3’ end of the double stranded DNA. [000424] In some embodiments, the molecular tags (e.g., barcode oligonucleotides) include unique molecular identifiers (UMIs). In some embodiments, the UMIs are oligonucleotides. In some embodiments, the molecular tags are coupled to priming sequences. In some embodiments, each of the priming sequences comprises a random N-mer sequence. In some embodiments, the random N-mer sequence is complementary to a 3’ sequence of the RNA molecules. In some embodiments, the priming sequence comprises a poly-dT sequence having a length of at least 5 bases. In some embodiments, the priming sequence comprises a poly-dT sequence having a length of at least 10 bases (SEQ ID NO: 4). In some embodiments, the priming sequence comprises a poly-dT sequence having a length of at least 5 bases, at least 6 bases, at least 7 bases, at least 8 bases, at least 9 bases, at least 10 bases. [000425] Unique molecular identifiers (UMIs), e.g., in the form of nucleic acid sequences are assigned or associated with individual cells or populations of cells, in order to tag or label the cell’s components (and as a result, its characteristics). These unique molecular identifiers may be used to attribute the cell’s components and characteristics to an individual cell or group of cells, additionally to be used as a method for counting the individual cells or groups of cells by their incorporation. [000426] In some aspects, the unique molecular identifiers are provided in the form of nucleic acid molecules (e.g., oligonucleotides) that comprise nucleic acid barcode sequences that may be attached to or otherwise associated with the nucleic acid contents of individual cell, or to other components of the cell, and particularly to fragments of those nucleic acids. The nucleic acid molecule can, and do have differing barcode sequences, or at least represent a large number of 115 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    different barcode sequences across all of the partitions in a given analysis. In some aspects, only one nucleic acid barcode sequence can be associated with a given partition, although in some cases, two or more different barcode sequences may be present. [000427] The nucleic acid barcode sequences can include from about 6 to about 20 or more nucleotides within the sequence of the nucleic acid molecules (e.g., oligonucleotides). The nucleic acid barcode sequences can include from about 6 to about 20, 30, 40, 50, 60, 70, 80, 90, 100 or more nucleotides. In some cases, the length of a barcode sequence may be about 6, 7, 8, 9, 10, 11, 12, 13, 14, 15, 16, 17, 18, 19, 20 nucleotides or longer. In some cases, the length of a barcode sequence may be at least about 6, 7, 8, 9, 10, 11, 12, 13, 14, 15, 16, 17, 18, 19, 20 nucleotides or longer. In some cases, the length of a barcode sequence may be at most about 6, 7, 8, 9, 10, 11, 12, 13, 14, 15, 16, 17, 18, 19, 20 nucleotides or shorter. These nucleotides may be completely contiguous, i.e., in a single stretch of adjacent nucleotides, or they may be separated into two or more separate subsequences that are separated by 1 or more nucleotides. In some cases, separated barcode subsequences can be from about 4 to about 16 nucleotides in length. In some cases, the barcode subsequence may be about 4, 5, 6, 7, 8, 9, 10, 11, 12, 13, 14, 15, 16 nucleotides or longer. In some cases, the barcode subsequence may be at least about 4, 5, 6, 7, 8, 9, 10, 11, 12, 13, 14, 15, 16 nucleotides or longer. In some cases, the barcode subsequence may be at most about 4, 5, 6, 7, 8, 9, 10, 11, 12, 13, 14, 15, 16 nucleotides or shorter. [000428] Moreover, when a population of barcodes is partitioned, the resulting population of partitions can also include a diverse barcode library that may include at least about 1,000 different barcode sequences, at least about 5,000 different barcode sequences, at least about 10,000 different barcode sequences, at least at least about 50,000 different barcode sequences, at least about 100,000 different barcode sequences, at least about 1,000,000 different barcode sequences, at least about 5,000,000 different barcode sequences, or at least about 10,000,000 different barcode sequences. Additionally, each partition of the population can include at least about 1,000 nucleic acid molecules, at least about 5,000 nucleic acid molecules, at least about 10,000 nucleic acid molecules, at least about 50,000 nucleic acid molecules, at least about 100,000 nucleic acid molecules, at least about 500,000 nucleic acids, at least about 1,000,000 nucleic acid molecules, at least about 5,000,000 nucleic acid molecules, at least about 10,000,000 nucleic acid molecules, at least about 50,000,000 nucleic acid molecules, at least 116 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    about 100,000,000 nucleic acid molecules, at least about 250,000,000 nucleic acid molecules and in some cases at least about 1 billion nucleic acid molecules. [000429] In some embodiments, the enhanced reverse transcriptase activity of the engineered reverse transcriptase disclosed herein is an enhanced ability to yield mitochondrial UMI counts as compared to a reverse transcriptase having the amino acid sequence set forth in SEQ ID NO:1 or 15. In some embodiments, the enhanced reverse transcriptase activity is an enhanced ability to yield increased ribosomal UMI counts as compared to a reverse transcriptase having the amino acid sequence set forth in SEQ ID NO:1 or 15. Read counting and UMI counting are the principal gene expression quantification schemes used in single-cell RNA-sequencing (scRNA- seq) analysis, as such with increased ribosomal UMI counts sensitivity and accuracy increases for a scRNA-seq assay in determining transcriptome profiles for any given cell, group of cells or tissues. Numerous metrics can be used for quality control of single-cell RNA-sequencing, including percent of reads mapping to ribosomal genes, percent of reads mapping to mitochondrial genes, total number of UMIs detected, or number of features to which 50% of the reads map. [000430] Beneficially, even following any subsequent amplification of the contents of a given partition, the number of different UMIs can be indicative of the quantity of mRNA originating from a given partition, and thus from the cell. As noted above, the transcripts can be amplified, purified and sequenced to identify the sequence of the cDNA transcript of the mRNA, as well as to sequence the barcode segment and the UMI segment. While a poly-dT primer sequence is described, other targeted or random primer sequences may also be used in priming the reverse transcription reaction. Likewise, although described as releasing the barcoded oligonucleotides into the partition, in some cases, the nucleic acid molecules bound to the bead (e.g., gel bead) may be used to hybridize and capture the mRNA on the solid phase of the bead, for example, in order to facilitate the separation of the RNA from other cell contents. [000431] It is recognized that certain reverse transcriptase enzymes may increase UMI reads from genes of a desired length or length of interest. The desired length of genes may be selected from lengths comprising less than 500 nucleotides, between 500 and 1000 nucleotides, between 1000 and 1500 nucleotides and greater than 1500 nucleotides. It is recognized that a reverse transcriptase may preferentially increase UMI reads from genes of one length range. It is 117 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    recognized that an engineered reverse transcriptase may perform similarly, differently or comparably in a 3’-reverse transcription assay or a 5’-reverse transcription assay. It is similarly recognized that an engineered reverse transcriptase may preferentially increase UMI reads from a length of genes in a 3’-reverse transcription assay than in a 5’-reverse transcription assay. 4. Gel bead [000432] The engineered reverse transcriptases or the recombinant RT protein of the present disclosure may be suitable for use in methods in which a cell can be co-partitioned along with a nucleic acid barcode molecule bearing bead. The nucleic acid barcode molecules can be released from the bead in the partition. By way of example, in the context of analyzing sample RNA, the poly-dT (poly-deoxythymine, also referred to as oligo (dT)) segment of one of the released nucleic acid molecules can hybridize to the poly-A tail of a mRNA molecule. Reverse transcription results in a cDNA transcript of the mRNA, but that transcript includes each of the sequence segments of the nucleic acid molecule. Without being limited by mechanism, because the nucleic acid molecule comprises an anchoring sequence, it may be more likely hybridize to and prime reverse transcription at the sequence end of the poly-A tail of the mRNA. Within any given partition, all of the cDNA transcripts of the individual mRNA molecules may include a common barcode sequence segment. However, the transcripts made from the different mRNA molecules within a given partition may vary at the unique UMI segment. [000433] Beneficially, even following any subsequent amplification of the contents of a given partition, the number of different UMIs can be indicative of the quantity of mRNA originating from a given partition, and thus from the cell. As noted above, the transcripts can be amplified, cleaned up and sequenced to identify the sequence of the cDNA transcript of the mRNA, as well as to sequence the barcode segment and the UMI segment. While a poly-dT primer sequence is described, other targeted or random priming sequences may also be used in priming the reverse transcription reaction. [000434] In some embodiments of the nucleic acid extension method described herein, the plurality of nucleic acid barcoded molecules are attached to a support (e.g., a particle, a slide, a chip, a bead, etc.). In one embodiment, the support is selected from an array, a bead, a gel bead, a microparticle, and a polymer. In some embodiments, the nucleic acid barcoded molecules attached to a support comprise molecular tags (UMIs), primer sequences, capture sequences, 118 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    cleavage sequences, or additional functional sequences. In some embodiments, the support is a gel bead. In that embodiment, the nucleic acid barcoded molecules are releasably attached to the gel bead. In some embodiments, the gel bead comprises a polyacrylamide polymer. [000435] In some embodiments, a cross-section of the gel bead is less than about 100 μm. In some embodiments, a cross-section of a gel bead is less than about 60 μm. In some embodiments, a cross-section of a gel bead is less than about 50 μm. In some embodiments, a cross-section of a gel bead is less than about 40 μm. In some embodiments, a cross-section of a gel bead is less than about 100 μm, less than about 99 μm, less than about 98 μm, less than about 97 μm, less than about 96 μm, less than about 95 μm, less than about 94 μm, less than about 93 μm, less than about 92 μm, less than about 91 μm, less than about 90 μm, less than about 89 μm, less than about 88 μm, less than about 87 μm, less than about 86 μm, less than about 85 μm, less than about 84 μm, less than about 83 μm, less than about 82 μm, less than about 81 μm, less than about 80 μm, less than about 79 μm, less than about 78 μm, less than about 77 μm, less than about 76 μm, less than about 75 μm, less than about 74 μm, less than about 73 μm, less than about 72 μm, less than about 71 μm, less than about 70 μm, less than about 69 μm, less than about 68 μm, less than about 67 μm, less than about 66 μm, less than about 65 μm, less than about 64 μm, less than about 63 μm, less than about 62 μm, less than about 61 μm, or less than about 60 μm. [000436] Functionalization of beads for attachment of nucleic acid molecules (e.g., oligonucleotides) may be achieved through a wide range of different approaches, including activation of chemical groups within a polymer, incorporation of active or activatable functional groups in the polymer structure, or attachment at the pre-polymer or monomer stage in bead production. [000437] For example, precursors (e.g., monomers, cross-linkers) that are polymerized to form a bead may comprise acrydite moieties, such that when a bead is generated, the bead also comprises acrydite moieties. The acrydite moieties can be attached to a nucleic acid molecule (e.g., oligonucleotide), which may include a priming sequence (e.g., a primer for amplifying target nucleic acids, random primer, primer sequence for messenger RNA) and/or one or more barcode sequences. The one more barcode sequences may include sequences that are the same for all nucleic acid molecules coupled to a given bead and/or sequences that are different across 119 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    all nucleic acid molecules coupled to the given bead. The nucleic acid molecule may be incorporated into the bead. [000438] In some cases, the nucleic acid molecule can comprise a functional sequence, for example, for attachment to a sequencing flow cell, such as, for example, a P5 sequence for Illumina® sequencing. In some cases, the nucleic acid molecule or derivative thereof (e.g., oligonucleotide or polynucleotide generated from the nucleic acid molecule) can comprise another functional sequence, such as, for example, a P7 sequence for attachment to a sequencing flow cell for Illumina® sequencing. In some cases, the nucleic acid molecule can comprise a barcode sequence. In some cases, the primer can comprise a unique molecular identifier (UMI). In some cases, the primer can comprise an R1 sequence for use in Illumina® sequencing workflows. In some cases, the primer can comprise an R2 sequence for use in Illumina® sequencing workflows. Examples of such nucleic acid molecules (e.g., oligonucleotides, polynucleotides, etc.) and uses thereof, as may be used with compositions, devices, methods and systems of the present disclosure, are provided in U.S. Patent Pub. Nos.2014/0378345 and 2015/0376609, each of which is entirely incorporated herein by reference. However, the present invention is not limited as to a composition of any nucleic acid molecule or derivative thereof, or any particular sequencing platform and these characterizations serve as examples only which may be useful in a reverse transcription workflow. [000439] In operation, a cell can be co-partitioned along with a barcode bearing bead. The barcoded nucleic acid molecules affixed to a bead can be released from the bead in the partition. By way of example, in the context of analyzing sample RNA, the poly-dT (poly-deoxythymine, also referred to as oligo (dT)) segment of one of the released nucleic acid molecules can hybridize to (e.g., capture)_the poly-A tail of a mRNA molecule. Reverse transcription may result in a cDNA transcript of the mRNA which cDNA transcript also includes each of the sequence segments of the nucleic acid molecule. Because the nucleic acid molecule comprises additional functional sequences (e.g., capture domains, primer domains, UMIs, barcodes, etc.), it can hybridize to and prime reverse transcription of the mRNA using the hybridized mRNA as a template. Within any given partition, all of the cDNA transcripts of the individual mRNA molecules may include a common barcode sequence. However, the transcripts made from the different mRNA molecules within a given partition may vary with respect to unique molecular 120 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    identifying sequences (e.g., UMIs). Beneficially, following any subsequent amplification of the contents of a given partition, the number of different UMIs can be indicative of the quantity of mRNA originating from a given partition, and thus from the cell. As noted above, the transcripts can be amplified and sequenced to identify the sequence of the original mRNA captured template, as well as the sequence of the associated barcode and UMI. While a poly-dT capture sequence is described, other targeted or random capture sequences may also be used in capture or hybridize to a template for initiating the reverse transcription reaction. [000440] In various embodiments, the poly-dT segment may be extended in a reverse transcription reaction using the mRNA as a template to produce a cDNA transcript complementary to the mRNA and also includes sequence segments of a barcode oligonucleotide. Terminal transferase activity of the reverse transcriptase can add additional bases to the cDNA transcript (e.g., polyC). The switch oligo may then hybridize with the additional bases added to the cDNA transcript and facilitate template switching. A sequence complementary to the switch oligo sequence can then be incorporated into the cDNA transcript via extension of the cDNA transcript using the switch oligo as a template. Within any given partition, all the cDNA transcripts of the individual mRNA molecules include a common barcode sequence segment. However, by including the unique random N-mer sequence, the transcripts made from different mRNA molecules within a given partition will vary at this unique sequence. As described elsewhere herein, this provides a quantification feature that can be identifiable even following any subsequent amplification of the contents of a given partition, e.g., the number of unique segments associated with a common barcode can be indicative of the quantity of mRNA originating from a single partition, and thus, a single cell. The cDNA transcript may then be amplified with PCR primers. The amplified product may then be purified (e.g., via solid phase reversible immobilization (SPRI)). The amplified product may be sheared, ligated to additional functional sequences, and further amplified (e.g., via PCR). [000441] Any of the engineered RT enzymes of the present disclosure, including without limitation any of the enzymes comprising the amino acid sequence and/or non-limiting embodiment of nucleic acid sequences shown in Table 1, or Table 2, could be analyzed in any suitable assay, including without limitation the assays described herein. Assays include without limitation 5’ gene expression analyses, with or without VDJ analysis, 3’ gene expression 121 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    analysis, epigenetic analysis, or multiomic analyses. In non-limiting embodiments, experiments are carried out as found in the manufacturer’s instructions for the Chromium Single Cell 5’ Gene Expression Assay kit (10X Genomics); Chromium Single Cell 3’ Gene Expression Assay kit (10X Genomics), including any of multiomic extensions or applications. C. Biological sample [000442] Methods disclosed herein can be performed on any type of sample. The sample can comprise a cell, a cell bead, a permeabilized cell, or a nucleus. The cell can be fixed. The cell can be permeabilized. Alternatively, the cell can be permeabilized and fixed. In some embodiments, the cell bead is fixed. In some embodiments, the nucleus is permeabilized or fixed. In some embodiments, the nucleus is permeabilized and fixed. In some embodiments, the biological sample comprises a suitable cellular preparation selected from cell populations and/or single cells. In some embodiments, the biological sample comprises a tissue. [000443] The sample can also be a suitable cellular preparation selected from cell populations and/or single cells. The sample can be a tissue. In some embodiments, the sample comprises cells in suspension, fresh cells, or fixed cells. The sample can also comprise cells and tissues immobilized on various solid surfaces. [000444] In some embodiments, when the biological sample is a cell, a cell bead, or a nucleus, the reverse transcription reaction described herein is part of a single cell RNA sequencing assay. In one embodiment, the single cell RNA sequencing assay further comprises, prior to the reverse transcription, partitioning the cell, the cell bead, or the nucleus into a partition. In one embodiment, the single cell RNA sequencing assay further comprises, after the reverse transcription reaction, hybridizing the nucleic acid product to an oligonucleotide molecule comprising a partition-specific barcode. [000445] In some embodiments, when the biological sample is a cell or a tissue sample immobilized on a surface, the reverse transcription reaction described herein is part of a spatial RNA sequencing assay. [000446] In some embodiments, the sample is a fresh tissue. In some embodiments, the sample is a frozen sample. In some embodiments, the sample was previously frozen. In some embodiments, the sample is a formalin-fixed, or paraffin embedded (FFPE) sample. FFPE 122 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    samples generally are heavily cross-linked and fragmented, and therefore this type of sample allows for limited RNA recovery using conventional detection techniques. In certain embodiments, methods of targeted RNA capture provided herein are less affected by RNA degradation associated with FFPE fixation than other methods (e.g., methods that take advantage of oligo-dT capture and reverse transcription of mRNA). In certain embodiments, methods provided herein enable sensitive measurement of specific genes of interest that otherwise might be missed with a whole transcriptomic approach. [000447] In some embodiments, a biological sample (e.g., tissue section) can be fixed with methanol, stained with hematoxylin and eosin, and imaged. In some embodiments, fixing, staining, and imaging occurs before one or more oligonucleotide probes are hybridized to the sample. Some embodiments of any of the workflows described herein can further include a destaining step (e.g., a hematoxylin and eosin destaining step), after imaging of the sample and prior to permeabilizing the sample. For example, destaining can be performed by performing one or more (e.g., one, two, three, four, or five) washing steps (e.g., one or more (e.g., one, two, three, four, or five) washing steps performed using a buffer including HCl). The images can be used to map spatial gene expression patterns back to the biological sample. A permeabilization enzyme can be used to permeabilize the biological sample directly on the slide. [000448] In some embodiments, the methods of targeted RNA capture as disclosed herein include hybridization of multiple probe oligonucleotides. In some embodiments, the methods include 2, 3, 4, or more probe oligonucleotides that hybridize to one or more analytes of interest. In some embodiments, the methods include two probe oligonucleotides. In some embodiments, the probe oligonucleotide includes sequences complementary that are complementary or substantially complementary to an analyte. For example, in some embodiments, the probe oligonucleotide includes a sequence that is complementary or substantially complementary to an analyte (e.g., an mRNA of interest (e.g., to a portion of the sequence of an mRNA of interest)). Methods provided herein may be applied to a single nucleic acid molecule or a plurality of nucleic acid molecules. A method of analyzing a sample comprising a nucleic acid molecule may comprise providing a plurality of nucleic acid molecules (e.g., RNA molecules), where each nucleic acid molecule comprises a first target region (e.g., a sequence that is 3′ of a target sequence or a sequence that is 5′ of a target sequence) and a second target region (e.g., a 123 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    sequence that is 5′ of a target sequence or a sequence that is 3′ of a target sequence), a plurality of first probe oligonucleotides, and a plurality of second probe oligonucleotides. [000449] In some embodiments, the templated ligation methods that allow for targeted RNA capture as provided herein include a first probe oligonucleotide and a second probe oligonucleotide. The first and second probe oligonucleotides each include sequences that are substantially complementary to the sequence of an analyte of interest. By substantially complementary, it is meant that the first and/or second probe oligonucleotide is at least 70%, at least 75%, at least 80%, at least 85%, at least 90%, at least 91%, at least 92%, at least 93%, at least 94%, at least 95%, at least 96%, at least 97%, at least 98%, at least 99%, or 100% complementary to a sequence in an analyte. In some instances, the first probe oligonucleotide and the second probe oligonucleotide hybridize to adjacent sequences on an analyte. [000450] In some embodiments, the first and/or second probe as disclosed herein includes one of at least two ribonucleic acid bases at the 3′ end; a functional sequence; a phosphorylated nucleotide at the 5′ end; and/or a capture probe binding domain. In some embodiments, the functional sequence is a primer sequence. The capture probe binding domain is a sequence that is complementary to a particular capture domain present in a capture probe. In some embodiments, the capture probe binding domain includes a poly(A) sequence. In some embodiments, the capture probe binding domain includes a poly-uridine sequence, a poly-thymidine sequence, or both. In some embodiments, the capture probe binding domain includes a random sequence (e.g., a random hexamer or octamer). In some embodiments, the capture probe binding domain is complementary to a capture domain in a capture probe that detects a particular target(s) of interest. [000451] In some embodiments, a capture probe binding domain blocking moiety that interacts with the capture probe binding domain is provided. In some instances, the capture probe binding domain blocking moiety includes a nucleic acid sequence. In some instances, the capture probe binding domain blocking moiety is a DNA oligonucleotide. In some instances, the capture probe binding domain blocking moiety is an RNA oligonucleotide. In some embodiments, a capture probe binding domain blocking moiety includes a sequence that is complementary or substantially complementary to a capture probe binding domain. In some embodiments, a capture probe binding domain blocking moiety prevents the capture probe binding domain from binding 124 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    the capture probe when present. In some embodiments, a capture probe binding domain blocking moiety is removed prior to binding the capture probe binding domain (e.g., present in a ligated probe) to a capture probe. In some embodiments, a capture probe binding domain blocking moiety comprises a poly-uridine sequence, a poly-thymidine sequence, or both. [000452] In some embodiments, the first probe oligonucleotide hybridizes to an analyte. In some embodiments, the second probe oligonucleotide hybridizes to an analyte. In some embodiments, both the first probe oligonucleotide and the second probe oligonucleotide hybridize to an analyte. Hybridization can occur at a target having a sequence that is 100% complementary to the probe oligonucleotide(s). In some embodiments, hybridization can occur at a target having a sequence that is at least (e.g., at least about) 80%, at least (e.g., at least about) 85%, at least (e.g., at least about) 90%, at least (e.g., at least about) 95%, at least (e.g., at least about) 96%, at least (e.g., at least about) 97%, at least (e.g., at least about) 98%, or at least (e.g., at least about) 99% complementary to the probe oligonucleotide(s). [000453] After hybridization of the first and second probe oligonucleotides, in some embodiments, the first probe oligonucleotide is extended. After hybridization, in some embodiments, the second probe oligonucleotide is extended. Extending probes can be accomplished using any method disclosed herein. In some instances, a polymerase (e.g., a DNA polymerase) extends the first and/or second oligonucleotide. [000454] In some embodiments, methods disclosed herein include a wash step. In some instances, the wash step occurs after hybridizing the first and the second probe oligonucleotides. The wash step removes any unbound oligonucleotides and can be performed using any technique or solution disclosed herein or known in the art. In some embodiments, multiple wash steps are performed to remove unbound oligonucleotides. [000455] In some embodiments, after hybridization of probe oligonucleotides (e.g., first and the second probe oligonucleotides) to the analyte, the probe oligonucleotides (e.g., the first probe oligonucleotide and the second probe oligonucleotide) are ligated together, creating a single ligated probe that is complementary to the analyte. Ligation can be performed enzymatically or chemically, as described herein. 125 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    [000456] In some embodiments of the method described herein, the engineered RT polypeptide or the recombinant RT protein enhances template switching (TS) efficiency, processivity efficiency, binding affinity, transcription efficiency, chemical tolerance, ability to yield mitochondrial unique molecular identity (UMI) counts, ability to yield ribosomal unique molecular identity (UMI) counts, strand displacement, end-to-end template jumping, or any combination thereof, when compared to an RT polypeptide or a recombinant RT protein lacking a conjugated DNA binding domain. [000457] In some embodiments, the engineered RT polypeptide or the recombinant RT protein enhances at least two or more of template switching (TS) efficiency, processivity efficiency, binding affinity, transcription efficiency, chemical tolerance, ability to yield mitochondrial unique molecular identity (UMI) counts, strand displacement, end-to-end template jumping, or ability to yield ribosomal unique molecular identity (UMI) counts, when compared to an RT polypeptide or a recombinant RT protein lacking a conjugated DNA binding domain. [000458] In some embodiments, the recombinant RT protein or the engineered RT enhances transcript capture during amplification, when compared to an RT polypeptide or a recombinant RT protein lacking a conjugated DNA binding domain. [000459] In some embodiments, the DNA binding domain of the recombinant RT protein or the engineered RT enhances the hybridization of a transcript and a primer during the amplification process, when compared to an RT polypeptide or a recombinant RT protein lacking a conjugated DNA binding domain.  D. Immunoprofiling [000460] Engineered reverse transcriptase polypeptides or recombinant RT proteins described herein may be used in methods of a T-Cell receptor (TCR) and a B-cell receptor (BRC) profiling. [000461] In some embodiments, an engineered reverse transcriptase is used in methods including but not limited to processing of a TCR from an individual T cell(s) or groups of T cell(s), determining the nucleotide sequence of the TCR(s) of T cell(s), and obtaining TCR repertoire profile. In some methods, a nucleic acid barcode sequence is appended to a nucleic acid molecule encoding for a TCR (e.g. a molecule derived from a T cell containing a nucleic 126 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    acid sequence encoding for a TCR, such as a TCRα and/or a TCRβ mRNA) resulting in a barcoded nucleic acid molecule comprising a sequence corresponding to a nucleic acid sequence of the TCR (e.g. comprises a V(D)J region of a TCR gene or a reverse complement thereof) and a sequence corresponding to the barcode sequence (which in some instances is the reverse complement of the barcode sequence present in the nucleic acid barcode molecule). A barcoded nucleic acid molecule may serve as a template, such as a template polynucleotide, that can be further processed (e.g., amplified) and sequenced to obtain the target nucleic acid sequence. For example, a barcoded nucleic acid molecule may be further processed (e.g., amplified) and sequenced to obtain the nucleic acid sequence of the TCR. [000462] TCR is a molecule found on the surface of T cells. Typically binding of the TCR by an antigenic molecule results in cell activation and response. The TCR is a heterodimer composed of two different protein chains. In many T cells, these two proteins are alpha (α) and beta (β) chains. In a smaller percentage of T cells, these two proteins are gamma (γ) and delta (δ) chains. The ratio of TCRs comprised of α/β chains versus γ/δ chains may change during a diseased state such as cancer, tumor, infectious disease, inflammatory disease or autoimmune disease. Engagement of the TCR with a peptide-MHC activates a T cell through a series of biochemical events mediated by associated enzymes, co-receptors, specialized adaptor molecules, and activated or released transcription factors. [000463] Each of the two chains of a TCR contains multiple copies of gene segments- a variable ‘V’ gene segment, a diversity ‘D’ segment and a joining ‘J’ segment. The TCR alpha chain is generated by recombination of V and J segments, while the beta chain is generated by recombination of V, D and J segments. Similarly, generation of the TCR gamma chain involves recombination of V and J segments. Generation of the TCR delta chain occurs by recombination of V, D and J gene segments. The intersection of these specific regions (V and J for the alpha or gamma chain, or V,D, J for the beta or delta chain) corresponds to the CDR3 region involved in antigen-MHC recognition. Complementarity determining regions (e.g., CDR1, CDR2 and CDR3) or hypervariable regions are sequences in the variable domains of antigen receptors (e.g., T cell receptor and immunoglobulin) that can complement an antigen. Most of the diversity of CDRs is found in CDR3, with the diversity being generated by somatic recombination events during the development of T lymphocytes. CDR3, which is encoded by the junctional region 127 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    between the V and J or D and J genes, is highly variable. CDR3 is often used as a region of interest to determine T cell clonotypes, a unique nucleotide sequence that arises during the gene rearrangement process, as it is highly unlikely that two T cells will express the same CDR3 nucleotide sequence unless they are derived from the same clonally expanded T cell. [000464] Because an active TCR consists of paired chains within single T cells, determination of the active paired chains within single T cells, determination of the active paired chains requires the sequencing of single T cells. TCR gene sequences may include, but are not limited to, sequences of various T cell receptor alpha variable genes (TRAV genes), T cell receptor alpha joining genes (TRAJ genes), T cell receptor alpha constant genes (TRAC genes), T cell receptor beta variable genes (TRBV genes), T cell receptor beta diversity genes (TRBD genes), T cell receptor beta joining genes (TRBJ genes), T cell receptor gamma variable genes (TRGV genes), T cell receptor gamma joining genes (TRGJ genes), T cell receptor gamma constant genes (TRGC genes), T cell receptor delta variable genes (TRDV genes), T cell receptor delta diversity genes (TRDD genes), T cell receptor delta joining genes (TRDJ genes) and T cell receptor delta constant genes (TRDC genes). VII. KITS [000465] One aspect of the present disclosure provides a kit comprising the engineered reverse transcriptase polypeptide or recombinant RT protein, the DNA binding domains or a derivative thereof as described herein. In some embodiments, the kit comprises one or more of a vector, a nucleotide, a buffer, a composition, a salt, and/or instructions. In another embodiment, a kit may comprise an engineered reverse transcriptase polypeptide or recombinant RT protein or a derivative thereof for use in reverse transcription or amplification of a nucleic acid molecule. In yet another embodiment, a kit may be used for single cell profiling of the transcriptome. In yet another embodiment, a kit may be used for spatial transcriptomics methods and assays. In yet another embodiment, a kit may be used for in situ methods and assays. [000466] The kit may include suitable reaction buffers, dNTPs, one or more primers, one or more control reagents, or any other reagents disclosed for performing the methods of the present disclosure. The engineered reverse transcriptase polypeptide or recombinant RT protein or a derivative thereof, reaction buffer, and dNTPs may be provided separately or may be provided together in a master mix solution. When the engineered reverse transcriptase polypeptide or 128 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    recombinant RT protein or a derivative thereof, reaction buffer, and dNTPs are provided in a master mix, the master mix is present at a concentration at least two times the working concentration indicated in instructions for use in an extension reaction. In other cases, the master mix may be present at a concentration at least three times, at least four times, at least five times, at least six times, at least seven times, at least eight times, at least nine times, or at least ten times, the working concentration indicated. The primer in the kits may be a poly-dT primer, a random N-mer primer, or a target-specific primer. [000467] The kits may further include one, two, three, four, five or more, up to all of partitioning fluids, including both aqueous buffers and non-aqueous partitioning fluids or oils, nucleic acid barcode capture probes that are releasably associated with beads, as described herein, microfluidic devices, reagents for disrupting cells, reagents for amplifying nucleic acids, as well as instructions for using any of the foregoing in the methods described herein. [000468] The instructions for using any of the methods are generally recorded on a suitable recording medium (e.g., printed on a substrate such as paper or plastic), or available in a digital format. As such, the instructions may be present in the kits as a package insert, in the labeling of the container of the kit or components thereof (i.e., associated with the packaging or subpackaging). In some cases, the instructions may be present as an electronic storage data file present on a suitable computer readable storage medium. In other cases, the actual instructions may not be present in the kit but means for obtaining the instructions from a remote source, e.g., via the internet, may be provided. For example, a kit that includes a web address where the instructions may be viewed and/or from which the instructions may be downloaded. As with the instructions, this means for obtaining the instructions is recorded on a suitable substrate. [000469] Kits according to this aspect of the present disclosure comprise a carrier means, such as a box, carton, tube or the like, having in close confinement therein one or more container means, such as vials, tubes, ampoules, bottles and the like, wherein a first container means contains one or more of the engineered reverse transcriptase polypeptide or recombinant RT proteins or derivatives thereof of the present disclosure having reverse transcriptase activity. When more than one polypeptide having reverse transcriptase activity is used, they may be in a single container as mixtures of two or more engineered reverse transcriptase polypeptide or recombinant RT proteins or derivatives thereof, or in separate containers. The kits of the 129 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    disclosure can also comprise (in the same or separate containers) one or more DNA polymerases, a suitable buffer, one or more nucleotides and/or one or more primers. [000470] The kits of the disclosure can also comprise one or more hosts or cells including those that are competent to take up nucleic acids (e.g., DNA molecules including vectors). Preferred hosts may include chemically competent or electrocompetent bacteria such as E. coli (including DH5, DH5α, DH10B, HB101, Top 10, and other K-12 strains as well as E. coli B and E. coli W strains). [000471] In a specific aspect of the present disclosure, the kits of the disclosure (e.g., reverse transcription and amplification kits) can include one or more components (in mixtures or separately) including one or more engineered reverse transcriptase polypeptides or recombinant RT proteins or derivatives thereof having reverse transcriptase activity of the disclosure, one or more nucleotides (one or more of which may be labeled, e.g., fluorescently labeled) used for synthesis of a nucleic acid molecule, and/or one or more primers (e.g., oligo(dT) for reverse transcription, randomers for extension reactions, etc.). Such kits can comprise one or more DNA polymerases. VIII. DEFINITIONS [000472] Unless defined otherwise, all technical and scientific terms used herein generally have the same meaning as commonly understood by one of ordinary skill in the art to which this technology belongs. As used in this specification and the appended claims, the singular forms “a”, “an” and “the” include plural referents unless the content clearly dictates otherwise. “A and/or B” is used herein to include all of the following alternatives: “A”, “B”, “A or B”, and “A and B”. For example, reference to “a cell” includes a combination of two or more cells, and the like. Generally, the nomenclature used herein and the laboratory procedures in cell culture, molecular genetics, organic chemistry, analytical chemistry and nucleic acid chemistry and hybridization described below are those well-known and commonly employed in the art. [000473] Where values are described as ranges, it will be understood that such disclosure includes the disclosure of all possible sub-ranges within such ranges, as well as specific numerical values that fall within such ranges irrespective of whether a specific numerical value or specific sub-range is expressly stated. 130 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    [000474] Whenever the term “at least,” “greater than,” or “greater than or equal to” precedes the first numerical value in a series of two or more numerical values, the term “at least,” “greater than” or “greater than or equal to” applies to each of the numerical values in that series of numerical values. For example, greater than or equal to 1, 2, or 3 is equivalent to greater than or equal to 1, greater than or equal to 2, or greater than or equal to 3. [000475] Whenever the term “no more than,” “less than,” or “less than or equal to” precedes the first numerical value in a series of two or more numerical values, the term “no more than,” “less than,” or “less than or equal to” applies to each of the numerical values in that series of numerical values. For example, less than or equal to 3, 2, or 1 is equivalent to less than or equal to 3, less than or equal to 2, or less than or equal to 1. [000476] Certain ranges are presented herein with numerical values being preceded by the term “about.” The term “About” is used herein to provide literal support for the exact number that it precedes, as well as a number that is near to or approximately the number that the term precedes. In determining whether a number is near to or approximately a specifically recited number, the near or approximating unrecited number may be a number which, in the context in which it is presented, provides the substantial equivalent of the specifically recited number. If the degree of approximation is not otherwise clear from the context, “about” means either within plus or minus 10% of the provided value, or rounded to the nearest significant figure, in all cases inclusive of the provided value. In some embodiments, the term “about” indicates the designated value ± up to 10%, up to ± 5%, or up to ± 1%. [000477] In some embodiments, the term “about” or “approximately” as used herein means within an acceptable error range for the particular value as determined by one of ordinary skill in the art, which will depend in part on how the value is measured or determined, i.e., the limitations of the measurement system. For example, “about” can mean within an acceptable standard deviation, per the practice in the art. Alternatively, “about” can mean a range of up to ±20%, preferably up to ±10%, more preferably up to ±5%, and more preferably still up to ±1% of a given value. Alternatively, particularly with respect to biological systems or processes, the term can mean within an order of magnitude, preferably within 2-fold, of a value. Where particular values are described in the application and claims, unless otherwise stated, the term 131 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    “about” is implicit and in this context means within an acceptable error range for the particular value. [000478] Numeric ranges are inclusive of the numbers defining the range. The term about is used herein to mean plus or minus ten percent (10%) of a value. For example, “about 100” refers to any number between 90 and 110. [000479] Headings, e.g., (a), (b), (i) etc., are presented merely for ease of reading the specification and claims. The use of headings in the specification or claims does not require the steps or elements be performed in alphabetical or numerical order or the order in which they are presented. [000480] Use of ordinal terms such as “first”, “second”, “third”, etc., in the claims to modify a claim element does not by itself connote any priority, precedence, or order of one claim element over another or the temporal order in which acts of a method are performed, but are used merely as labels to distinguish one claim element having a certain name from another element having a same name (but for use of the ordinal term) to distinguish the claim elements. Similarly, the use of these terms in the specification does not by itself connote any required priority, precedence, or order. [000481] As used herein, the term “Analyte” is intended a biological molecule. Analytes include but are not limited to a DNA analyte, an RNA analyte, an oligonucleotide, a reporter molecule, a reporter molecule configured to directly couple to a protein, a reporter molecule configured to indirectly couple to a protein, a reporter molecule configured to directly couple to a metabolite, and a reporter molecule configured to indirectly couple to a metabolite. [000482] The terms “Adaptor(s),” “Adapter(s)” and “Tag(s)” may be used synonymously. An adaptor or tag can be coupled to a polynucleotide sequence to be “tagged” by any approach, including ligation, hybridization, or other approaches. [000483] As used herein, the term “Barcoded nucleic acid molecule” generally refers to a nucleic acid molecule that results from, for example, the processing of a nucleic acid barcoded molecule with a nucleic acid sequence (e.g., nucleic acid sequence complementary to a nucleic acid primer sequence encompassed by the nucleic acid barcoded molecule). The nucleic acid sequence may be a targeted sequence or a non-targeted sequence. The nucleic acid barcoded 132 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    molecule may be coupled to or attached to the nucleic acid molecule comprising the nucleic acid sequence. For example, a nucleic acid barcoded molecule described herein may be hybridized to an analyte (e.g., a messenger RNA (mRNA) molecule) of a cell. Reverse transcription can generate a barcoded nucleic acid molecule that has a sequence corresponding to the nucleic acid sequence of the mRNA and the barcode sequence (or a reverse complement thereof). The processing of the nucleic acid molecule comprising the nucleic acid sequence, the nucleic acid barcoded molecule, or both, can include a nucleic acid reaction, such as, in non-limiting examples, reverse transcription, nucleic acid extension, ligation, etc. The nucleic acid reaction may be performed prior to, during, or following barcoding of the nucleic acid sequence to generate the barcoded nucleic acid molecule. For example, the nucleic acid molecule comprising the nucleic acid sequence may be subjected to reverse transcription and then be attached to the nucleic acid barcoded molecule to generate the barcoded nucleic acid molecule, or the nucleic acid molecule comprising the nucleic acid sequence may be attached to the nucleic acid barcoded molecule and subjected to a nucleic acid reaction (e.g., extension, ligation) to generate the barcoded nucleic acid molecule. A barcoded nucleic acid molecule may serve as a template, such as a template polynucleotide, that can be further processed (e.g., amplified) and sequenced to obtain the target nucleic acid sequence. For example, in the methods and systems described herein, a barcoded nucleic acid molecule may be further processed (e.g., amplified) and sequenced to obtain the nucleic acid sequence of the nucleic acid molecule (e.g., mRNA). [000484] A nucleic acid barcoded molecule of a plurality of nucleic acid molecules may be used to generate a “barcoded nucleic acid molecule.” In some cases, a barcoded molecule comprises a different reporter barcode sequence that identifies a second analyte. A different reporter barcode sequence or an analyte-specific barcode sequence may identify a protein, a lipid, a metabolite or other second analyte. [000485] Barcoded nucleic acids may be generated (e.g., via a nucleic acid reaction, such as nucleic acid extension or ligation) from the constructs described in FIG.12. For example, capture handle sequence may then be hybridized to complementary sequence, such as capture sequence 1223 to generate (e.g., via a nucleic acid reaction, such as nucleic acid extension or ligation) a barcoded nucleic acid molecule comprising cell (e.g., partition specific) barcode sequence 1222 (or a reverse complement thereof) and reporter barcode sequence 1222 (or a 133 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    reverse complement thereof). In some embodiments capture handle sequence 4323 comprises a sequence complementary to a template switching oligonucleotide on the capture sequence 1223. In some embodiments, the nucleic acid barcoded molecule 1290 (e.g., partition-specific barcoded molecule) further includes a UMI (not shown). Barcoded nucleic acid molecules can then be optionally processed as described elsewhere herein, e.g., to amplify the molecules and/or append sequencing platform specific sequences to the fragments. See, e.g., U.S. Pat. Pub.2018/0105808, which is hereby entirely incorporated by reference for all purposes. Barcoded nucleic acid molecules, or derivatives generated therefrom, can then be sequenced on a suitable sequencing platform. [000486] In some instances, analysis of multiple analytes (e.g., nucleic acids and one or more analytes using labelling agents described herein) may be performed. In some instances, analysis of an analyte (e.g., a nucleic acid, a polypeptide, a carbohydrate, a lipid, a glycan, a glycan motif, a metabolite, a protein, etc.) comprises a workflow as generally depicted in FIG.12. A nucleic acid barcoded molecule 1290 (e.g., partition specific barcoded molecule) may be co-partitioned with the one or more analytes. In some instances, nucleic acid barcoded molecule 1290 is attached to a support 1230 (e.g., a bead, such as a gel bead), such as those described elsewhere herein. For example, nucleic acid barcoded molecule 1290 may be attached to support 1230 via a releasable linkage 1240 (e.g., comprising a labile bond), such as those described elsewhere herein. Nucleic acid barcoded molecule 1290 may comprise a functional sequence 1221 and optionally comprise other additional sequences, for example, a barcode sequence 1222 (e.g., common barcode, partition-specific barcode, or other functional sequences described elsewhere herein), and/or a UMI sequence (not shown). The nucleic acid barcoded molecule 1290 may comprise a capture sequence 1223 that may be complementary to another nucleic acid sequence, such that it may hybridize to a particular sequence, e.g., capture handle sequence 1223. [000487] For example, capture sequence 1223 may comprise a poly-T sequence and may be used to hybridize to mRNA. Referring to FIG.12, in some embodiments, nucleic acid barcoded molecule 1290 comprises capture sequence 1223 complementary to a sequence of RNA molecule 1260 from a cell. In some instances, capture sequence 1223 comprises a sequence specific for an RNA molecule. Capture sequence 1223 may comprise a known or targeted sequence or a random sequence. In some instances, a nucleic acid extension reaction may be 134 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    performed, thereby generating a barcoded nucleic acid product comprising capture sequence 12323, the functional sequence 1221, barcode sequence 1222, any other functional sequence, and a sequence corresponding to the RNA molecule 1260. [000488] In another example, capture sequence 1223 may be complementary to an overhang sequence or an adapter sequence that has been appended to an analyte. Any suitable agent may degrade beads. Suitable agents may include, but are not limited to, changes in temperature, changes in pH, reduction, oxidation and exposure to water or other aqueous solutions. [000489] In some instances, a cell that is bound to labelling agent which is conjugated to oligonucleotide and support 1230 (e.g., a bead, such as a gel bead) comprising nucleic acid barcoded molecule 1290 is partitioned into a partition amongst a plurality of partitions (e.g., a droplet of a droplet emulsion, a well of a microwell array, a fixed cell and/or nucleus, a fixed and permeabilized cell and/or nucleus). [000490] The term “Bead,” as used herein, generally refers to a particle. The bead may be a solid or semi-solid particle. The bead may be a gel bead. The gel bead may include a polymer matrix (e.g., matrix formed by polymerization or cross-linking). The polymer matrix may include one or more polymers (e.g., polymers having different functional groups or repeat units). Polymers in the polymer matrix may be randomly arranged, such as in random copolymers, and/or have ordered structures, such as in block copolymers. Cross-linking can be via covalent, ionic, or inductive, interactions, or physical entanglement. The bead may be a macromolecule. The bead may be formed of nucleic acid molecules bound together. The bead may be formed via covalent or non-covalent assembly of molecules (e.g., macromolecules), such as monomers or polymers. Such polymers or monomers may be natural or synthetic. Such polymers or monomers may be or include, for example, nucleic acid molecules (e.g., DNA or RNA). The bead may be formed of a polymeric material. The bead may be magnetic or non-magnetic. The bead may be rigid. The bead may be flexible and/or compressible. The bead may be disruptable or dissolvable. The bead may be a solid particle (e.g., a metal-based particle including but not limited to iron oxide, gold or silver) covered with a coating comprising one or more polymers. Such coating may be disruptable or dissolvable. [000491] As used herein, the term “Each,” when used in reference to a collection of items, is intended to identify an individual item in the collection but does not necessarily refer to every 135 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    item in the collection, unless expressly stated otherwise, or unless the context of the usage clearly indicates otherwise. [000492] As used herein, the term “Efficiency” in the context of a nucleic acid modifying enzyme of this invention refers to the ability of the enzyme to perform its catalytic function under specific reaction conditions. Typically, “efficiency” as defined herein is indicated by the amount of product generated under given reaction conditions. [000493] As used herein, the term “Enhances” in the context of an enzyme refers to improving the activity of the enzyme, i.e., increasing the amount of product per unit enzyme per unit time. [000494] As used herein, the term “Fidelity” refers to the accuracy of polymerization, or the ability of the reverse transcriptase to discriminate correct from incorrect substrates, (e.g., nucleotides) when synthesizing nucleic acid molecules which are complementary to a template. The higher the fidelity of a reverse transcriptase, the less the reverse transcriptase misincorporates nucleotides in the growing strand during nucleic acid synthesis; that is, an increase or enhancement in fidelity results in a more faithful reverse transcriptase having decreased error rate or decreased misincorporation rate. [000495] As used herein, the term "% homology," which is used interchangeably with the term "% identity," refers to the level of nucleic acid or amino acid sequence identity between the nucleic acid sequence that encodes any one of the inventive polypeptides (e.g., variant reverse transcriptases) or the inventive polypeptide's amino acid sequence, when aligned using a sequence alignment program. [000496] As used herein, the term “Identical” in the context of two nucleic acids or polypeptide sequences refers to the residues in the two sequences that are the same when aligned for maximum correspondence, as measured using a sequence comparison algorithms. Sequence comparison algorithms are known to those skill in the art. See. e.g., ebi.ac.uk/Tools/msa/clustalo/. [000497] As used herein, the term “Inhibitor resistance” refers to the ability of a reverse transcriptase to perform reverse transcription in the presence of a compound, chemical, protein, buffer, etc. that is typically inhibitory to the reverse transcriptase (prevents or inhibits reverse transcriptase activity). 136 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    [000498] As used herein, the term “Low volume reaction” means a reaction volume less than 1 nanoliter, less than 750 picoliters, or less than 500 picoliters. [000499] The term “Molecular tag,” as used herein, generally refers to a molecule capable of binding to a macromolecular constituent. The molecular tag may bind to the macromolecular constituent with high affinity. The molecular tag may bind to the macromolecular constituent with high specificity. The molecular tag may comprise a nucleotide sequence. The molecular tag may comprise a nucleic acid sequence. The nucleic acid sequence may be at least a portion or an entirety of the molecular tag. The molecular tag may be a nucleic acid molecule or may be part of a nucleic acid molecule. The molecular tag may be an oligonucleotide or a polypeptide. The molecular tag may comprise a DNA aptamer. The molecular tag may be or comprise a primer. The molecular tag may be, or comprise, a protein. The molecular tag may comprise a polypeptide. The molecular tag may be a barcode. [000500] As used herein, the term “mutation” or “mutant” or “variant“ indicates a change or changes introduced in a wild-type DNA sequence or a wildtype amino acid sequence. Examples of mutations or variants include, but are not limited to, substitutions, insertions, deletions, and point mutations. Mutations can be made either at the nucleic acid level or at the amino acid level. [000501] As used herein, the term “Operably linked” or “conjugated” or “fusion” means that, in relation to the engineered RT polypeptide or the recombinant RT protein sequence, there are one or more sequences at the N or C terminus that, when transcribed and translated, create additional polypeptides in association with the enzyme amino acid sequence, thereby created a conjugation or fusion of one or more polypeptides from one expression vector. [000502] The term “Partition,” as used herein, generally, refers to a space or volume that may be suitable to contain one or more species or conduct one or more reactions. A partition may be a physical compartment, such as a droplet, well or a fixed and/or permeabilized cell and/or nucleus. The partition may isolate space or volume from another space or volume. The droplet may be a first phase (e.g., aqueous phase) in a second phase (e.g., oil) immiscible with the first phase. The droplet may be a first phase in a second phase that does not phase separate from the first phase, such as, for example, a capsule or liposome in an aqueous phase. A partition may comprise one or more other (inner) partitions. In some cases, a partition may be a virtual compartment that can be defined and identified by an index (e.g., indexed libraries) across 137 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    multiple and/or remote physical compartments. For example, a physical compartment may comprise a plurality of virtual compartments. [000503] The term “Partitioning” as used herein is intended to encompass parting, dividing, depositing, separating, or compartmentalizing into one or more partitions. Systems and methods for partitioning of one or more particles (such as, but not limited to, biological particles, macromolecular constituents of biological particles, beads, reagents, etc.) into discrete compartments or partitions (referred to interchangeably here as partitions), wherein each partition maintains separation of its own content from the contents of other partitions are known in the art. See for example US 2020/0032335, herein incorporated by reference in its entirety. The partition can be a droplet in an emulsion. A partition may comprise one or more other partitions. [000504] A “plurality of nucleic acid barcoded molecules” may comprise at least about 500 nucleic acid barcoded molecules, at least about 1,000 nucleic acid barcoded molecules, at least about 5,000 nucleic acid barcoded molecules, at least about 10,000 nucleic acid barcoded molecules, at least about 50,000 nucleic acid barcoded molecules, at least about 100,000 nucleic acid barcoded molecules, at least about 500,000 nucleic acid barcoded molecules, at least about 1,000,000 barcoded molecules, at least about 5,000,000 nucleic acid barcoded molecules, at least about 10,000,000 nucleic acid barcoded molecules, at least about 100,000,000 nucleic acid barcoded molecules, at least about 1,000,000,000 nucleic acid barcoded molecules. In some cases, a plurality of nucleic acid barcoded molecules comprise a partition-specific barcode sequence. [000505] Each of the plurality of nucleic acid barcoded molecules may include an identifier sequence separate from the partition-specific barcode sequence, where the identifier sequence is different for each nucleic acid partition-specific barcoded molecule of the plurality of nucleic acid partition specific barcoded molecules. In some cases, such an identifier sequence is a unique molecular identifier (UMI) as described elsewhere herein. As described elsewhere herein, UMI sequences can uniquely identify a particular nucleic acid molecule that is barcoded, which may be identifying particular nucleic acid molecules that are analyzed, counting particular nucleic acid molecules that are analyzed, etc. Furthermore, in some cases, each of the plurality of nucleic acid barcoded molecules can comprise the partition specific barcode sequence and the bead can 138 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    be from plurality of beads, such as a population of barcoded beads. Each of the partition specific barcode sequences can be different from partition specific barcode sequences of nucleic acid barcoded molecules of other beads of the plurality of beads. Where this is the case, a population of barcoded beads, with each bead comprising a different partition specific barcode sequence can be analyzed. [000506] As used herein, the term “Processivity” refers to the ability of a reverse transcriptase to continuously extend a primer without disassociating from the nucleic acid template. The length of a template a reverse transcriptase or polymerase is capable of replicating can also be used to describe the processivity of that reverse transcriptase or polymerase. In some embodiments, “Processivity” refers to the ability of a polymerase to remain bound to the template or substrate and perform DNA synthesis. Processivity is measured by the number of catalytic events that take place per binding event. [000507] As used herein, “Purified” means that a molecule is present in a sample at a concentration of at least 95% by weight, or at least 98% by weight of the sample in which it is contained. [000508] As used herein, the term “Reverse transcriptase activity,” “reverse transcription activity,” or “reverse transcription” indicates the capability of an enzyme to synthesize a DNA strand (that is, complementary DNA or cDNA) using RNA as a template. Reverse transcriptase activity may be measured by incubating an enzyme in the presence of an RNA template and deoxynucleotides, in the presence of an appropriate buffer, under appropriate conditions, for example as described in the Example below. Methods for measuring RT activity are provided in the example below and also are well known in the art. Bosworth, et al., Nature 1989, 341:167- 168. [000509] As used herein, the term “recombinant RT” comprises the engineered RT fusion protein described herein or the engineered RT variant described herein. [000510] As used herein, the term “Reverse transcriptase (RT)” is used in its broadest sense to refer to any enzyme that exhibits reverse transcription activity as measured by methods disclosed herein or known in the art. A "reverse transcriptase" of the present invention, therefore, includes reverse transcriptases from retroviruses, other viruses, as well as a DNA polymerase exhibiting 139 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    reverse transcriptase activity, such as Tth DNA polymerase, Taq DNA polymerase, Tne DNA polymerase, Tma DNA polymerase, etc. RT from retroviruses include, but are not limited to, Moloney Murine Leukemia Virus (M-MLV) RT, Human Immunodeficiency Virus (HIV) RT, Avian Sarcoma-Leukosis Virus (ASLV) RT, Rous Sarcoma Virus (RSV) RT, Avian Myeloblastosis Virus (AMV) RT, Avian Erythroblastosis Virus (AEV) Helper Virus MCAV RT, Avian Myelocytomatosis Virus MC29 Helper Virus MCAV RT, Avian Reticuloendotheliosis Virus (REV-T) Helper Virus REV-A RT, Avian Sarcoma Virus UR2 Helper Virus UR2AV RT, Avian Sarcoma Virus Y73 Helper Virus YAV RT, Rous Associated Virus (RAV) RT, and Myeloblastosis Associated Virus (MAV) RT, and as described in U.S. Patent Application 2003/0198944 (hereby incorporated by reference in its entirety). For review, see e.g., Levin, 1997, Cell, 88:5-8; Brosius et al.51995, Virus Genes 11:163-79. Known reverse transcriptases from viruses require a primer to synthesize a DNA transcript from an RNA template. Reverse transcriptase has been used primarily to transcribe RNA into cDNA, which can then be cloned into a vector for further manipulation or used in various amplification methods such as polymerase chain reaction (PCR), nucleic acid sequence-based amplification (NASBA), transcription mediated amplification (TMA), or self-sustained sequence replication (3SR). [000511] The term “Sample,” as used herein, generally refers to a biological sample of a subject. The biological sample may comprise any number of macromolecules, for example, cellular macromolecules. The sample may be a cell sample. The sample may be a cell line or cell culture sample. The sample can include one or more cells. The sample can include one or more microbes. The biological sample may be a nucleic acid sample or protein sample. The biological sample may also be a carbohydrate sample or a lipid sample. The biological sample may be derived from another sample. The sample may be a tissue sample, such as a biopsy, core biopsy, needle aspirate, or fine needle aspirate. The sample may be a fluid sample, such as a blood sample, urine sample, or saliva sample. The sample may be a skin sample. The sample may be a cheek swab. The sample may be a plasma or serum sample. The sample may be a cell-free or cell free sample. A cell-free sample may include extracellular polynucleotides. Extracellular polynucleotides may be isolated from a bodily sample that may be selected from blood, plasma, serum, urine, saliva, mucosal excretions, sputum, stool and tears. 140 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    [000512] As used herein, the term “Sequencing,” generally refers to methods and technologies for determining the sequence of nucleotide bases in one or more polynucleotides. Any method of sequencing known in the art may be used to evaluate the products of a reaction performed by an engineered reverse transcriptase of the current application. Sequencing can be performed by various systems currently available, such as, without limitation, a sequencing system by Illumina®, Pacific Biosciences (PacBio®), Oxford Nanopore®, or Life Technologies (Ion Torrent®). Alternatively, or in addition, sequencing may be performed using nucleic acid amplification, polymerase chain reaction (PCR) (e.g., digital PCR, quantitative PCR, or real time PCR), or isothermal amplification. In some examples, such systems provide sequencing reads (also “reads” herein). A read may include a string of nucleic acid bases corresponding to a sequence of a nucleic acid molecule that has been sequenced. In some situations, systems and methods provided herein may be used with proteomic information. [000513] As used herein, the term “Substantially complementary” means that a first sequence is at least 50%, 55%, 60%, 65%, 70%, 75%, 80%, 85%, 90%, 95%, 97%, 98% or 99% identical to the complement of a second sequence over a region of 2, 3, 4, 5, 6, 7, 8, 9, 10, 11, 12, 13, 14, 15, 16, 17, 18, 19, 20-40, 40-60, 60-100, or more nucleotides, or that the two sequences hybridize under stringent hybridization conditions. Substantially complementary also means that a sequence in one strand is not completely and/or perfectly complementary to a sequence in an opposing strand, but that sufficient bonding occurs between bases on the two strands to form a stable hybrid complex in set of hybridization conditions (e.g., salt concentration and temperature). Such conditions can be predicted by using the sequences and standard mathematical calculations known to those skilled in the art. [000514] The term “Subject,” as used herein, generally refers to an animal, such as a mammal (e.g., human) or avian (e.g., bird), or other organism, such as a plant. For example, the subject can be a vertebrate, a mammal, a rodent (e.g., a mouse), a primate, a simian or a human. Animals may include, but are not limited to, farm animals, sport animals, and pets. A subject can be a healthy or asymptomatic individual, an individual that has or is suspected of having a disease (e.g., cancer) or a pre-disposition to the disease, and/or an individual that is in need of therapy or suspected of needing therapy. A subject can be a patient. A subject can be a microorganism or microbe (e.g., bacteria, fungi, archaea, viruses). 141 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    [000515] As used herein, the term “Thermoreactivity” or “Thermoreactive” refers to the ability of a reverse transcriptase to exhibit enzyme activity at elevated temperatures. [000516] As used herein, “Thermostability” or “thermostable” refers to the ability of a reverse transcriptase to withstand exposure to elevated temperatures, but not necessarily show activity at such elevated temperatures. In some embodiments, a thermostable reverse transcriptase or polymerase refers to any enzyme that catalyzes polynucleotide synthesis by addition of nucleotide units to a nucleotide chain using DNA or RNA as a template and has an optimal activity at a temperature above 53° C. [000517] As used herein, the terms “Unique molecular identifier”, “Unique molecular identifying sequence”, “UMI” and “UMI sequence” are used synonymously. Individual barcoded molecules may comprise a common barcode sequence such as a partition specific sequence or a spatial array where every capture probe has a unique barcode sequence. [000518] By “Binding sequence” is intended a nucleic acid sequence capable of binding to an analyte. [000519] As used herein, the term “Variant” means a protein which is derived from a precursor protein (such as the native protein, for example MMLV native protein as set forth in SEQ ID NO:7) by addition of one or more amino acids to either or both the C- and N-terminal end, substitution of one or more amino acids at one or a number of different sites in the amino acid sequence, deletion of one or more amino acids at either or both ends of the protein or at one or more sites in the amino acid sequence, or addition of a fusion domain. SEQ ID NO:1 is a variant of MMLV. The preparation of an enzyme variant is preferably achieved by modifying a DNA sequence which encodes for the wild-type protein, transformation of that DNA sequence into a suitable host, and expression of the modified DNA sequence to form the derivative enzyme. A variant reverse transcriptase of the invention includes a protein comprising altered amino acid sequences in comparison with a precursor enzyme amino acid sequence wherein the variant reverse transcriptase retains the characteristic enzymatic nature of the precursor enzyme but which may have altered properties in some specific aspect. For example, an engineered reverse transcriptase variant may have an altered pH optimum or increased temperature stability but may retain its characteristic transcriptase activity. 142 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    [000520] A “Variant” may have at least about 45%, at least about 50%, at least about 55%, at least about 60%, at least about 65%, at least about 70%, at least about 75%, at least about 80%, at least about 85%, at least about 88%, at least about 90%, at least about 91 %, at least about 92%, at least about 93%, at least about 94%, at least about 95%, at least about 96%, at least about 97%, at least about 98%, at least about 99%, or at least about 99.5% sequence identity to a polypeptide sequence when optimally aligned for comparison. Percent identity may pertain to the percent identity of the DNA binding domain or the engineered reverse transcriptase portion of an engineered reverse transcriptase. As used herein, a variant residue position is described in relation to the wild-type or precursor amino acid sequence set forth in SEQ ID NO:7; the amino acid position is indexed to SEQ ID NO:7. A fusion variant comprises at least one fusion domain selected from DNA binding domains described elsewhere herein. [000521] As used herein, a protein having a certain percent (e.g., at least 90%, at least 91%, at least 92%, at least 93%, at least 94%, at least 95%, at least 96%, at least 97%, at least 98%, or at least 99%) of sequence identity with another sequence means that, when aligned, that percentage of bases or amino acid residues are the same in comparing the two sequences. This alignment and the percent homology or identity can be determined using any suitable software program known in the art, for example those described in CURRENT PROTOCOLS IN MOLECULAR BIOLOGY, Ausubel et al., eds., 1987, Supplement 30, section 7.7.18. Representative programs include the Vector NTI Advance™ 9.0 (Invitrogen Corp. Carlsbad, CA), GCG Pileup, FASTA (Pearson et al. (1988) Proc. Natl Acad. ScL USA 85:2444-2448), and BLAST (BLAST Manual, Altschul et al., Nat’l Cent. Biotechnol. Inf., Nat’l Lib. Med. (NCIB NLM NIH), Bethesda, Md., and Altschul et al., (1997) Nucleic Acids Res.25:3389-3402) programs. Another typical alignment program is ALIGN Plus (Scientific and Educational Software, PA), generally using default parameters. Other sequence alignment software programs that find use are the TFASTA Data Searching Program available in the Sequence Software Package Version 6.0 (Genetics Computer Group, University of Wisconsin, Madison, WI and CLC Main Workbench (Qiagen) Version 20.0. The present disclosure is not limited to the software being used to align two or more sequences. [000522] As used herein, the term “Wild-type” or “WT” refers to a gene or gene product that has the characteristics of that gene or gene product when isolated from a naturally occurring 143 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    source. The amino acid sequence set forth in SEQ ID NO:7 is a WT Murine Moloney Leukemia Virus (MMLV) sequence (Genbank NP_955591.1 p80 RT). [000523] Unless otherwise indicated, nucleic acids are written left to right in 5' to 3' orientation; amino acid sequences are written left to right in amino to carboxy orientation, respectively. [000524] The headings provided herein are not limitations of the various aspects or embodiments of the invention which can be had by reference to the specification as a whole. Accordingly, the terms defined immediately below are more fully defined by reference to the specification as a whole. EXAMPLES [000525] It will be understood that the reference to the below examples is for illustration purposes only and do not limit the scope of the claims. Each aspect, embodiment, or feature of the invention may be combined with any other aspect, embodiment, or feature the invention unless clearly indicated to the contrary. Unless defined otherwise, all technical and scientific terms used herein have the meaning commonly understood by a person skilled in the art to which this invention belongs. [000526] Exemplary engineered RTs comprising a MMLV RT variant (42BL) operably linked N-terminally to a DNA binding protein derived from a budding yeast DAT1 were generated. However, engineered RT comprising C-terminal fusions can also be constructed. Dat from different budding yeasts that contains at least three repeated pentads of G-R-K-P-G can be used. An DAT1 N-terminal fusion protein was generated; and an DAT1C-terminal fusion protein was generated. The DAT1 fusion proteins are produced with an N-terminal 6x His Tag and thrombin cleavage site. The 6x His Tag was used for purification purposes and removed by thrombin cleavage. [000527] Exemplary engineered RTs (e.g., variant MMLV) comprising an RT operably linked N-terminally to a truncated DAT1 molecule comprising 90 amino acid (DAT(90) was also generated. While the majority of exemplary engineered RT tested below are of N-terminal fusions, C-terminal RT fusions can also be constructed. In fact, a non-exhaustive list of possible 144 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    fusion proteins can be generated. These different reverse transcriptases can also be fused to full length DAT protein, DAT1 truncated to 36 amino acids, or DAT1(90) homologs. [000528] In vitro characterizations of an exemplary engineered RT (N-Dat_42BL) were conducted as shown below. These results showed the enhanced performance of engineered reverse transcriptases comprising DAT1 described herein in a capillary electrophoresis (CE) assay (FIG.13); a Single Cell 5’ (SC-5’) or a Single Cell 3’ (SC-3’) gene expression assay (FIGs.14A-B); or Visium ATP reverse transcriptase based assay (FIGs.15A-B). [000529] It was demonstrated that DAT1 can be truncated to a minimal binding domain of 36 amino acids as shown by N-Dat36_42BL and C-Dat36_42BL. In addition, homologs of DAT1 from other organisms can also be used to generate the engineered reverse transcriptase described herein as shown by N-DAT1-TL-QID04042BL; N-DAT1-TL-XP36142BL; N-DAT1-TL- XP55842BL; or N-DAT1-TL-XP68342BL. [000530] DAT1(90) can be fused to other reverse transcriptase enzymes, such as other variants of MMLV reverse transcriptases as shown by N-DAT-9042B; N-DAT-9050A+ G; N-DAT-90 SOLD 33 VDG; N-DAT-90 SOLD 01; C-DAT-90 SOLD 01. Example 1: Capillary Electrophoresis Assay Validation [000531] Reverse transcription and sequencing reactions were prepared. The reaction volume was 50 µl and reactions contained 5’-end labeled FAM Reverse Transcriptase primer 2, RT Reagent B (Chromium Next GEM Single Cell Reagent, 10X Genomics), RNA template (RNA Temp 2), template switching oligo 1 (TSO1), and the indicated engineered reverse transcriptase. Table 3A: Capillary Electrophoresis Assay Reactants R k Fi l
Figure imgf000146_0001
145 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    Table 3A: Capillary Electrophoresis Assay Reactants Reagent Stock Final on
Figure imgf000147_0001
5’ kit (10X Genomics, Inc), except the reverse transcriptase was altered for a particular reaction. Stock concentrations and final concentrations in the reactions are shown in Tables 3A-B. Variations of the assay stock concentrations and final concentrations in the reactions shown in Table 4 were used. The reactions included stoichiometrically equal amounts of enzyme and template for single turnover conditions. [000533] Additionally, reverse transcription and sequencing reactions were prepared using GAPDH or GRCh38 as a template. The reaction volume was 50 µl; reactions contained 5’-end labeled GAPDH or GRCh38 Primer, GEM-U reagent, RNA template (GAPDH or GRCh38 template, template switching oligo 1 (TSO1, and the indicated engineered reverse transcriptase. Stock concentrations and final concentrations in the reactions are shown in Table 3B. The reactions included stoichiometrically equal amounts of enzyme and template for single turnover conditions. Reactants were incubated at 53°C for 45 minutes, then diluted 1:20 in HiDi formamide. The formamide mixture was heated to 95°C for 5 mins, then chilled on ice for 2 mins. Samples were loaded on the CE, the DS-33 dye set was selected and long fragment analysis was performed using the GS1200LIZ size standard. The GEM-U reagent approximates the formulation of the actual reagent mixture in a GEM assay when the contents of the Z1 and Z2 channels are mixed. Results from one such experiment are shown in FIGs.22-26. [000534] Reactants were incubated at 53°C for one hour, then diluted 1:40 in water and then 1:20 in HiDi formamide. The formamide mixture was heated to 95°C for 5 mins, then chilled on ice for 2 mins. Samples were loaded on a Seqstudio™ (Thermo Fisher Scientific) and fragment analysis by capillary electrophoresis was carried out with the appropriate dye channels and size standards. The assay was validated with synthetically sized oligonucleotides and with a transcription positive, template switching null engineered reverse transcriptase and a transcription positive, template switching positive reverse transcriptase (Enzyme Mix C,). The GEM-U reagent approximates the formulation of the actual reagent mixture in a GEM assay 146 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    when the contents of the Z1 and Z2 channels are mixed. Capillary Electrophoresis Assay Reactants are disclosed in Table 1A, Capillary Electrophoresis Assay template, Primer and TSO sequences are shown in Table 4A. [000535] Reverse transcription and sequencing reactions were also prepared using GAPDH or GRCh38 as a template. The reaction volume was 50 µl; reactions contained 5’-end labeled GAPDH or GRCh38 primer, GEM-U reagent, RNA template (GAPDH or GRCh38 template), template switching oligo 1 (TSO1), and the indicated engineered reverse transcriptase. Stock concentrations and final concentrations in the reactions are shown in Table 4B. The reactions included stoichiometrically equal amounts of enzyme and template for single turnover conditions. Reactants were incubated at 53°C for 45 minutes, then diluted 1:20 in HiDi formamide. The formamide mixture was heated to 95°C for 5 mins, then chilled on ice for 2 mins. Samples were loaded on the CE, the DS-33 dye set was selected and long fragment analysis was performed using the GS1200LIZ size standard. The GEM-U reagent approximates the formulation of the actual reagent mixture in a GEM assay when the contents of the Z1 and Z2 channels are mixed. [000536] Results from one such experiment are shown in FIGs 22-26. [000537] In particular, the reaction volume was 50 µl; reactions contained 5’-end labeled GAPDH primer, GEM-U reagent, RNA template (GAPDH template), template switching oligo 1 (TSO1), and the indicated engineered reverse transcriptase(s). The final concentrations in the reactions are shown in Table 4B. The reaction buffer was SOP for SC-5’ and the reaction time was 45 minutes. [000538] Tables 4A-B show Capillary Electrophoresis (CE) Assay Reactants and Template, Primer and TSO sequences (SEQ ID NOS:173, 175, 176, respectively in order of appearance.) Table 4A: Capillary Electrophoresis Assay Template, Primer and TSO sequences
Figure imgf000148_0001
147 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    Table 4A: Capillary Electrophoresis Assay Template, Primer and TSO sequences Reagent Stock Final Volume
Figure imgf000149_0001
[000539] Several mutants were constructed using a Q5 mutagenesis kit (NEB) with mutagenic primers per manufacturing instructions. Linearized products were circularized by KLD treatment (kinase, ligase, DpN1) and cloned. Several mutants were synthesized as whole plasmids and furnished by Twist Biosciences, South San Francisco CA. [000540] Briefly, a vector comprising the DAT1 sequence was obtained from Integrated DNA Technologies (IDT, Coralville, IA). Cloning was performed using a Gibson Assembly kit from New England Biolabs (NEB, Ipswitch, ME). Q5 polymerase was used to generate Gibson vectors. Amplification conditions were an initial denaturation at 95°C for 2.5 minutes, 30 cycles of denature (95°C, 30 sec), a 45 sec gradient annealing and extension at 72°C for 6 minutes, 35 sec, followed by a final extension at 72°C for 2 minutes. Amplification reactions with multiple annealing gradient temperatures (65.2°C, 67°C, 68.5°C and 69.6°C) were performed. [000541] Amplification products were evaluated on a 1.2% agarose E-Gel using SYBR-Safe. Products were pooled prior to clean-up. Cloning and expression were performed in the Acella cell line from EdgeBio (San Jose, CA). Cells were selected on LB-Kanamycin plates. DAT1 N- terminal and C-terminal fusions to an engineered reverse transcriptase having the amino acid sequence set forth in SEQ ID NO:1 were obtained by screening of bacterial colonies. The sequences of the fusion proteins were confirmed using method known in the art. 148 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    Example 3: Single Cell Sensitivity and Mapping [000542] Various engineered reverse transcriptases were evaluated in single cell experiments with human peripheral blood monocytes (PBMCs) and mouse peripheral blood monocyte cells (C57B/L6) using 3’ and 5’ reaction conditions as described above herein. Sensitivity and mapping were evaluated. Results from engineered reverse transcriptase (N-Dat 42BL (SEQ ID NO: 175) and C-DAT 42B L; SEQ ID NO: 174) were compared to results obtained from a commercially available engineered MMLV or MMLV variants (SOLD 001 (SEQ ID NO: 65); and SOLD 33 VDG (SEQ ID NO: 173). [000543] Results from one such series of experiments are summarized in FIG.20. The percent change is as compared to a commercially available variant MMLV reverse transcriptase. The change in median genes and median UMI’s queried at 20,000 reads per cell and the change in reads mapped to the transcriptome and reads mapped to exons are shown. A commercially available engineered reverse transcriptase was used as the control. Improvements in both the 5’ and 3’ chemistries were more pronounced in the mouse PBMC’s than in the human PBMCs. [000544] FIG.20 and FIG.21 show that DAT1 in combination with reverse transcriptase (MMLV, 42B, or other 42B variant thereof (e.g., SOLD 001 or SOLD 033 VDG)), either fused at the N-terminus or C-terminus of the RT improved GEX sensitivity even at low sequencing depth. Improvement was observed with both gene expression, which was increased by up to ~37% and UMI , which was increased by up to 13% as captured at 20k rrpc. Gains were even more significant at higher read depth. In addition, the engineered RT disclosed herein exhibited large change in differential gene expression in single cell assays. The engineered RT molecules comprising DAT1 or variant thereof disclosed herein picked-up to about 5000 additional genes when compared to a non-DAT1 RT (e.g., 42B). FIGs.20-26. These engineered RT polypeptide also exhibited increase in median UMI counts per spot and median gene counts per spot in spatial assay. Moreover, the engineered RT molecules gave decrease in fraction of reads mapped to exons with gain in fraction mapped to introns. See e.g., FIG.21. A performance difference between FPLC and plate purified proteins was performed. 149 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    Example 4: Analysis of engineered RT comprising a DAT1 DNA binding domain in 5’ single cell assay [000545] This example shows that an engineered reverse transcriptase described herein (e.g., 42B-L-dat fusion) significantly improved the sensitivity of a 5’ single cell assay. [000546] To determine the effectiveness of the engineered RT and/or recombinant RT protein described herein, the RT enzymes were analyzed in 5’ gene expression assay. Analysis of 42B- NDAT, 42B L-CDAT, 42B L-NDAT and 50A+G-NDATin 5’ single cell assay was performed and compared to an assay using 42B, 42B L, or 50A+G RT. All RT enzymes were used at 1.31 uM concentration, and all purified by the same method. RT enzymes tested included 42B (SEQ ID NO: 1, SEQ ID NO: 143, or SEQ ID NO: 172), 50A+G (Table 2; SEQ ID NO: 147), 42B_L (Table 2; SEQ ID NO: 145). [000547] FIGs.22A-B show the relative differences in performance of the engineered RT polypeptides or control non-DAT RT compared to control RT (42B). The median genes and UMIs/cell at 50k raw-reads per cell were used to compare the sensitivity of three reverse transcriptases with and without the DAT fusion domain.42B-NDAT showed 43.29% (Median genes/cell) and 39.80% (median UMIs/cell) enhancement over 42B alone.42BL-CDAT showed 30.07% (Median genes/cell) and 23.76% (median UMIs/cell) enhancement over 42B alone. 42BL-NDAT showed 47.40% (Median genes/cell) and 41.32% (median UMIs/cell) enhancement over 42B alone.50A+G-NDAT showed 45.00% (Median genes/cell) and 40.63% (median UMIs/cell) enhancement over 42B alone. In the same assay, non-DAT RT, 42 B L only showed 7.25% (Median genes/cell) and 13.91% (median UMIs/cell) enhancement over 42B alone. The non-DAT RT, 50A+G only showed 28.66% (Median genes/cell) and 38.01% (median UMIs/cell) enhancement over 42B alone. [000548] Furthermore, both a C-terminal fusion and an N-terminal fusion of DAT to 42BL increased median genes per cell and median UMIs per cell in the single cell assays. The C- terminal fusion increased median genes per cell by 21% compared to 42BL (3032 versus 2500) and median UMIs per cell by 9 % (8976 versus 8262) as compared to 42BL. The N-terminal fusion increased median genes per cell by 37% (3436 versus 2500) and median UMIs per cell by 24%. Additionally, the 50A+G-DAT fusion showed a 13 % increase in genes per cell (3380 versus 2999) and a 2% increase in UMIs per cell (10200 versus 10010). Thus, each DAT RT 150 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    fusion tested demonstrated improved quality metrics in a transcriptomics assay as compared to a matched control including the same RT but lacking the DAT domain. [000549] These results showed clear performance gains in engineered RT polypeptides comprising a DAT fusion on either the N-terminal or C-terminal domain. The results also demonstrated the enhanced sensitivity of a single cell assay using an engineered reverse transcriptase polypeptide or an engineered recombinant reverse transcriptase comprising DAT1 described herein as shown by the median genes identified per cell (Median genes/cell) or the median UMIs identified per cell (median UMIs/cell). Specifically, an engineered RT of the present disclosure significantly improved the RT sensitivity when compared to a non-DAT1 RT; and an engineered reverse transcriptase comprising a DAT1 binding domain further significantly increased the gain in sensitivity of the engineered reverse transcriptase described herein. [000550] The performance of various engineered RT disclosed herein were also analyzed at maximum normalization depth using the median genes/cell (FIG.23A) or the median UMIs/Cell (FIG.23B) to compare the generated library complexity of three reverse transcriptases (42B, 42B L, and 50A+G) with (N-DAT or C-DAT) and without the DAT domain. FIGs.23 C-D show saturation curves of the median genes (FIG.23C) and counts/cell (FIG.27C) as a function of read depth, which further demonstrate that the median genes and counts/cell were higher using the engineered RT with a DAT DNA binding domain when compared to MMLV variants lacking the DAT DNA binding domain. FIG.23 demonstrates a clear benefit of using the DAT DNA binding domain in a Single Cell 5’ (SC-5’) gene expression assay. [000551] To further determine the properties of the engineered RTs described herein, these RT enzymes were analyzed in a 5’ gene expression assay. All RT enzymes were used at 1.31 uM concentration, and all purified by the same method. [000552] A good correlation in differential gene expression (e.g., gene calling) between engineered RT variants comprising the DAT DNA binding domain) relative to non-DAT control was observed (FIGs.24-25). FIGs.24A-F and FIGs.25A-F showed significant levels of differential gene expression with 42B L-CDAT, 42B L-NDAT, 50A+G-NDAT. FIGs.24A-F shows the differential gene expression of some engineered RT comprising DAT1 at the N- terminus. FIGs.24A, C, and E feature scatter plots showing gene expression correlation of three reverse transcriptases with and without the DAT fusion domain. FIGs.24B, D, and F feature 151 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    volcano plots showing the number of differentially expressed genes between three reverse transcriptases with and without the DAT fusion domain. [000553] FIGs.25A-F illustrate the differential gene expression of some engineered RT comprising DAT1 at the C-terminus. FIGs.25A, C, and E feature scatter plots showing gene expression correlation comparison of engineered RT comprising N- Terminal DAT Domain and C- Terminal DAT Domain. FIGs.25B, D, and F feature volcano plots comparing the number of differentially expressed genes between various RT. Gains in performance, were still present, but not as significant with a C-terminal fusion. For example, the C-DAT terminal fusion did not perform as well as the N-terminal DAT fusion, but both N-terminal DAT fusion and C-DAT terminal fusion showed performance gains when compared to the RT backbone (42B L; SEQ ID NO: 145) alone. As the conditions become more similar in complexity (42B L - NDAT > 42B L - CDAT > 42B L, see FIGs 22-23), there are lower levels of DEGs and the feature scatter plots begin to correlate better. More genes will begin to pass the differential gene expression threshold (expression >= 0.3 and p-value <= 0.05) as the complexity gap widens, but it is not a 1:1 correlation. [000554] FIGs.26A-D show graphs illustrating the performance comparison of the impact of the DAT fusion domain across three reverse transcriptase backbones based on median genes (FIG.26A) and UMIs/cell at maximum normalization depth (FIG.26B), gene expression correlation (FIG.26C), and differential gene expression (FIG.26D). The aggregated metrics comparing the performance among 42B, 42B L, and 50A+G backbones with and without the DAT fusion showed a clear performance benefit from the DAT domain, such as e.g., enhanced sensitivity. Therefore, the 42B L-DAT1 fusion RT significantly improve in single cell assay performance. Example 5: Assays for analyses of Engineered RT polypeptides [000555] Any of the engineered RT enzymes of the invention, including without limitation any of the enzymes described in Table 1, or Table 2 could be analyzed in any suitable assay, including without limitation the assays described herein. Assays include without limitation 5’ gene expression analyses, with or without VDJ analysis, 3’ gene expression analysis, epigenetic analysis, or multiomic analyses. In non-limiting embodiments, experiments are carried out as found in the manufacturer’s instructions for the Chromium Single Cell 5’ Gene Expression 152 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    Assay kit (10X Genomics); Chromium Single Cell 3’ Gene Expression Assay kit (10X Genomics), including any of multiomic extensions or applications. Example 6: Single Cell 3’ and 5’ cDNA Yields [000556] Various engineered reverse transcriptases can be evaluated in single cell experiments with peripheral blood monocytes (PBMCs) at a cell load of 1,000, using the 3’ and 5’ configurations. Emulsion droplets can contain gel beads with either barcoded poly-dT primer sequences (3’ configuration) or barcoded template switch oligo sequences (5’ configuration) that also include a UMI and Illumina® Read 1 sequence. When cells are lysed within the droplet, the poly-dT primer hybridizes to the poly-A tail of the cellular mRNA, which is extended by the reverse transcriptase. Once the end of the template is reached, the reverse transcriptase will exhibit terminal transferase activity to add an overhang of three non-templated deoxycytidines (CCC) to the 3’ end of the synthesized cDNA. The CCC overhang will hybridize to the 3 riboguanosines (rGrGrG) present on the 3’ end of the template switch oligo, allowing the reverse transcriptase to “switch” templates and continue synthesis to the 5’ end of the template switch oligo. Depending on the which configuration of gel bead is used (3’ or 5’) the barcode and UMI will allow either the 3’ or 5’-end of the mRNA molecule to be identified in the final sequencing library. Following reverse transcription at 48 °C or 53 °C for 45 mins, and a 5 min heat-kill at 85 °C, droplets were broken and the cDNA was purified with Dynabeads. The cDNA was then amplified via PCR, purified with a 0.6x SPRI, and quantified with an Agilent Bioanalyzer using the DNA High Sensitivity Kit. The cDNA yield (ng) was then obtained. Example 7: Single Cell 3’ Quality Metrics [000557] Various engineered reverse transcriptases can be evaluated in single cell experiments with peripheral blood monocytes (PBMCs) using the 3’ and 5’ reaction conditions. Either 10 µL of the amplified cDNA (3’ conditions) or 20 µL containing a maximum of 50 ng of amplified cDNA (5’ conditions) can then be fragmented and A-tailed, cleaned with a double-sided SPRI (0.6x/0.8x), ligated to functional adaptors with an Illumina® Read 2 sequence, cleaned with a 0.8x SPRI, and then can be further amplified with sample indexing primers that include the P5 and P7 priming sites and the i5 and i7 sample indexes. The amplification product can be cleaned up with a double-side (0.6x/0.8x) SPRI, and the average size can be determined with an Agilent Bioanalyzer using the DNA High Sensitivity Kit. The material can then be quantified by qPCR 153 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    and pooled for next generation sequencing on an Illumina® Novaseq targeting a sequencing depth of at least 50,000 reads per cell and using the following run parameters (Read 1: 28 cycles, i7 Index: 10 cycles, i5 Index: 10 cycles, Read 2: 90 cycles). Data can be collected, demultiplexed, and processed. Standard quality metrics were obtained. [000558] Generally, the single cell 5’ reactions use less enzyme and TSO oligo than the single cell 3’ reactions. The 5’ TSO oligo is also twice the length of the 3’ TSO oligo with varied sequence context due to the presence of the UMI and the barcode. The single cell 5’ reaction conditions are generally considered a more stringent test of performance than the 3’ single cell reaction conditions. [000559] It is expected that the results from these experiments will show that the engineered reverse transcriptase polypeptide disclosed herein will show improved sensitivity at20 or 50 kilo reads per cell (krpc). [000560] Most of the engineered RT polypeptides will yield metrics within parity for valid UMI’s, valid barcodes, ribosomal UMI’s, mitochondrial UMI’s, transcript coverage, and reads with any poly A sequence, reads with any switch oligo sequence and reads with primer or homopolymer sequence. [000561] FIG.21 summarizes results from a series of experiments using the 5’ reaction conditions. The figure summarizes metrics of the 5’ single cell experiments, including 20k read metrics, 50K read metrics and reads mapped to the transcriptome. The engineered reverse transcriptase variants have the amino acid sequences provided set forth in SEQ ID NO: 65 (SOLD 001), SEQ ID NO: 173 (SOLD 33 VDG), SEQ ID NO: 174 (C-DAT 42BL), and SEQ ID NO: 175 (N-DAT 42BL). The percent indicates the percent change from the results obtained with an engineered reverse transcriptase having the amino acid sequence set forth in SEQ ID NO:1 or 179 (42B). In particular, the engineered reverse transcriptase polypeptides showed a significant improvement in sensitivity. [000562] FIG.21 shows additional metrics related to results obtained from the indicated engineered reverse transcriptase in single cell 5’ experiments. Most of the variants yielded metrics within parity for valid UMI’s, valid barcodes, ribosomal UMI’s, mitochondrial UMI’s, 154 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    transcript coverage, reads with any poly A sequence, reads with any switch oligo sequence and reads with primer or homopolymer sequence. Example 8: Immunoprofiling and TCR Improvements [000563] Immune profiling is an extension of the 5’ chemistry to profile genes specifically for T-cell and/or B-cell receptors in the mRNA pool. Methods of immune profiling are known in the art and generally include additional rounds of PCR on the cDNA with a pool of sequence specific primers to allow for targeted enrichment of T-cell and/or B-cell receptor genes. Immune profiling assays may also detect UMIs for B-cell receptor genes, namely IGH, IGK, and IGL (Immunoglobulin heavy chain (IGH), kappa (IGK), and light (IGL) chain). Immune profiling data is informative for immunology research and is an extension of standard gene expression evaluation. Methods of immune profiling include but are not limited to Chromium Next Gen Single Cell™ kits (10X Genomics, Pleasanton CA). [000564] Amplified cDNA (2 µl) from the 5’ configuration of reverse transcription reactions can be subjected to two additional rounds of PCR enrichment with TCR immune profiling, which included a double-sided (0.5x/0.8x) SPRI clean-up between the first and second round of thermal cycling reactions. The amplified products can then be cleaned-up with a subsequent double-sided (0.5x/0.8x) SPRI, fragmented and A-tailed, ligated to functional adaptors with an Illumina® Read 2 sequence, cleaned up with a 0.8x SPRI, and then further amplified with sample indexing primers that include the P5 and P7 priming sites and the i5 and i7 sample indexes. The amplification product can be cleaned up with a 0.8x SPRI, and average size can be determined with an Agilent Bioanalyzer using the DNA High Sensitivity Kit. The material can then be quantified by qPCR and can be pooled for next generation sequencing on an Illumina® Novaseq targeting a sequencing depth of at least 5,000 reads per cell and using the following run parameters (Read 1: 28 cycles, i7 Index: 10 cycles, i5 Index: 10 cycles, Read 2: 90 cycles). Data can be collected, demultiplexed, and single-cell V(D)J analysis can be performed. Results that are obtained from engineered reverse transcriptases can be compared to results are obtained from a commercially available enzyme or an RT lacking the DAT 1 DNA binding domain. The percent change in median TRA UMI’s and median TRB UMI’s from mouse and human PBMCs for each RT tested can be shown as a percent change in median IGH, IGK and IGL from mouse PBMC’s. It is expected that the median TRA UMIs and median TRB UMIs obtained with any 155 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    engineered reverse transcriptase described herein h will be greater than those obtained with a commercially available reverse transcriptase or non-DAT RT in both human PBMCs and mouse PBMCs. Engineered reverse transcriptases will also exhibit IG sensitivity that is either comparable or improved IG sensitivity as compared to previous results. Example 9: Spatial assays [000565] This example shows various methods for performing spatial analysis using the engineered RT enzymes of the present disclosure. The Example provides an exemplary method for detecting an individual gene expression or globally-expressed RNA transcripts in a fresh or frozen tissue using an engineered RT of SEQ ID NO: 175, which includes N-Terminal fusion of the first 90a.a. of Dat 1 (“N-DAT”) fused to a 42B variant of SEQ ID NO: 1, 143, 145, or 172) compared to a control RT variant (having the same RT enzyme but lacking the N-DAT or SEQ ID NO: _1, 143, 145, or 172) using three different reverse transcription buffer formulations, the commercially available RT reagent, buffer X.4 and buffer X.5. The commercially available RT reagent comprises 2% Glycerol, 50 mM Tris pH 8.3, 3 mM MgCl2, 75.04 mM KCl, 0.5% Supersonic F-108, 0.489 mg/ml BSA, and 3.16 mM dNTPs. Buffer X.4 and buffer X.5 have the same components as the commercial RT reagent with the following differences. Buffer X.4 contains 1.05% Glycerol, 158.3 mM NaCl, 0.579 mg/ml BSA, 4.685 mM dNTPs, 1.5mM dCTP, and 0.775mM GTP. Buffer X.5 contains 133.8 mM NaCl, 1% Supersonic F-108, 3.96 mM dNTPs, 1.31 mM dCTP, and 0.655 mM GTP. [000566] Sample preparation. A formalin-fixed, paraffin-embedded (FFPE) human tonsil sample, a frozen Human tonsil sample, or a fresh Human tonsil sample can be used for the method described herein. Human tonsil FFPE sections on standard slides (for sandwich conditions; FIGs.1A-B, FIGs.2A-B, and FIGs.3A-C) or gene expression (GEx) slides (for non-sandwich control conditions) were deparaffinized, H&E stained, and imaged. After the imaging, the human tonsil sections were hematoxylin-destained with HCL solution. The sections were then decrosslinked by incubating at 70°C for 1 hour in decrosslinking solution. Decrosslinking solution was removed and the tissues were incubation in 1x PBS-Tween for 15 minutes. [000567] Spatial analysis of the prepared fresh human tonsil sample or the human tonsil FFPE sections can be analyzed using the 3’ workflow shown in FIG.27. 156 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    [000568] Sample permeabilization. The human tonsil sample or the human tonsil FFPE sections were incubated with a permeabilization solution at 37°C for30 minutes. The permeabilization solution was washed out and the human tonsil sample was prepared for analyte capture by adding 0.1X SSC buffer and subjected to a pre-equilibration thermocycling protocol (e.g., lid temperature and pre-equilibrate at 53 °C, reverse transcription at 53 °C for 45 minutes, and then hold at 4 °C). The SSC buffer was removed. [000569] cDNA generation. A Master Mix comprising a commercially available RT reagent buffer, buffer X.4, or buffer X.5, nuclease-free water, a template switch oligo, a reducing agent, and a reverse transcriptase (control 42B RT, N-DAT-42B RT, or small scale-purified N-DAT- 42B RT) was added to the human tonsil sample and subjected to a thermocycling protocol, which comprised performing a reverse transcription reaction at 53 °C for 45 minutes and hold at 4 °C. A second strand synthesis was performed on the sample by subjecting the sample to a thermocycling protocol, which comprised e.g., pre-equilibrating at 65 °C, synthesizing the second strand at 65 °C for 15 minutes, then holding at 4 °C. The Master Mix reagents were removed from the sample and 0.8M KOH was added and incubated for 5 minutes at room temperature. The KOH was removed and followed by the addition of an elution buffer. [000570] A Second Strand Mix, including a second strand reagent, a second strand primer, and a second strand-polymerizing enzyme, can be added to the sample and the sample can be sealed and incubated. At the end of the incubation, the reagents can be removed and elution buffer can be added and removed from the sample, and 0.8 M KOH can be added again to the sample and the sample can be incubated for 10 minutes at room temperature. Tris-HCl can be added and the reagents can be mixed. The sample can be transferred to a new tube, vortexed, and placed on ice. [000571] cDNA amplification and quality control. A qPCR Mix, including nuclease-free water, qPCR Master Mix, and cDNA primers, was prepared and pipetted into wells in a qPCR plate. A small amount of the human tonsil sample was added to the plated qPCR Mix, and thermocycled according to a predetermined thermocycling protocol (e.g., step 1: 98 °C for 3 minutes, step 2: 98 °C for 5 seconds, step 3: 63 °C for 30 seconds, step 4: record amplification signal, step 5: repeating 98 °C for 5 seconds, 63 °C for 30 seconds for a total of 25 cycles). After completing the thermocycling, a cDNA amplification mix, including amplification mix and cDNA primers was prepared and combined with the remaining sample and mixed. The sample was incubated 157 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    and thermocycled (e.g., lid temperature at 105 °C for -45-60 minutes; step 1: 98 °C for 3 minutes, step 2: 98 °C for 15 seconds, step 3: 63 °C for 20 seconds, step 4: 72 °C for one minute, step 5: [the number of cycles determined by qPCR Cq Values], step 6: 72 °C for 1 minute, and step 7: hold at 4 °C). The sample can then be stored at 4 °C for up to 72 hours or at -20 °C for up to 1 week or use immediately. [000572] SPRI Cleanup. The cDNA sample was resuspended in 0.6X SPRI select Reagent and pipetted to ensure proper mixing. The sample was incubated for 5 minutes at room temperature and cleared by placing the sample on a magnet (e.g., the magnet is in the high position). The supernatant was removed and 80% ethanol was added to the pellet and incubated for 30 seconds. The ethanol was removed and the pellet washed. The sample was centrifuged and placed on a magnet (e.g., the magnet is on the low position). Any remaining ethanol was removed and the sample was air dried for up to 2 minutes. The magnet was removed and elution buffer was added to the sample, mixed, and incubated for 2 minutes at room temperature. The sample was then placed on the magnet (e.g., on low position) until the solution cleared. The sample was transferred to a new tube strip and stored at 4 °C for up to 72 hours or at -20 °C for up to 4 weeks. [000573] A portion of the sample was run on an Agilent Bioanalyzer High Sensitivity chip, where a region was selected and the cDNA concentration was measured to calculate the total cDNA yield. Alternatively, the quantification can be determined by Agilent Bioanalyzer or Agilent TapeStation. [000574] Fragmentation, End-repair, and A-tailing. Following SPRI Cleanup, the sample was processed
Figure imgf000159_0001
expression library construction. A Fragmentation Mix, including a fragmentation buffer and a fragmentation enzyme was prepared on ice. Elution buffer and fragmentation mix was added to each sample, mixed, and centrifuged. The sample mix was then placed in a thermocycler and cycled according to a predetermined protocol (e.g., lid temperature at 65 °C for ~ 35 minutes, pre-cool block down to 4 °C before fragmentation at 32 °C for 5 minutes, End-repair and A-tailing at 65 °C for 30 minutes and holding at 4 °C). [000575] SPRI Cleanup. The 0.6X SPRI select Reagent was added to the sample and incubated for 5 minutes at room temperature. The sample was placed on a magnet (e.g., in the high position) until the solution cleared, and the supernatant was transferred to a new tube strip.0.8X 158 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    SPRI select Reagent was added to the sample, mixed, and incubated for 5 minutes at room temperature. The sample was placed on a magnet (e.g., in the high position) until the solution cleared. The supernatant was removed and 80% ethanol was added to the pellet, incubated for 30 seconds, and washed. The ethanol wash was repeated and the sample placed on a magnet (e.g., in the low position) until the solution cleared. The remaining ethanol was removed and elution buffer was added to the sample, mixed, and incubated for 2 minutes at room temperature. The sample was placed on a magnet (e.g., in the high position) until the solution cleared, and a portion of the sample was moved to a new tube strip. [000576] Ligation. An Adaptor Ligation Mix, including ligation buffer, DNA ligase, and adaptor oligos, was prepared and centrifuged. The Adaptor Ligation Mix was added to the sample, pipette-mixed, and centrifuged briefly. The sample was then thermocycled using a predetermined protocol (e.g., lid temperature at 30 °C for ~15 minutes, step 1: 20 °C for 15 minutes, step 2: 4 °C hold). The sample was vortexed to resuspend SPRI select Reagent for another SPRI cleanup as described above. [000577] Sample Index PCR. An amplification mix was prepared and added to the sample. An individual Dual Index TT Set A was added to the sample, pipette-mixed and subjected to a pre- determined thermocycling protocol (e.g., lid temperature at 105 °C for -25-40 minutes, step 1: 98 °C for 45 seconds, step 2: 98 °C for 20 seconds, step 3: 54 °C for 30 seconds; step 4: 72 °C for 20 seconds, step 5: reverting to step 2 for a predetermined number of cycles, step 6: 72 °C for 1 minute, and 4 °C on hold). Vortex to resuspend the SPRI select Reagent for cleaning. [000578] The cleaned sample was stored at 4 °C for up to 72 hours, or at -20 °C for long-term storage. The average fragment size was determined using a Bioanalyzer trace or an Agilent TapeStation. The library was sequenced using a sequencing platform, for example, MiSeq, NextSeq 500/550, HiSeq 2500, HiSeq 3000/4000, NovaSeq, and iSeq. See, Illumina®, Indexed Sequencing Overview Guides, February 2018, Document 15057455v04; and Illumina® Adapter Sequences, May 2019, Document #1000000002694vl 1, each of which is hereby incorporated by reference, for information on P5, P7, i7, i5, TruSeq™ Read 2, indexed sequencing, and other reagents described herein. 159 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    Example 10: N-DAT RT variant improved the sensitivity of the spatial assay while maintaining good spatial resolution in human samples. [000579] To determine the improvement associated with the engineered N-DAT-42B RT variant, the effect of the 42B RT control variant (SEQ ID NO: 1, 143, 145, or 172) and the N- DAT-42B RT variant were compared using the three buffer types as previously described. The human tonsil samples were permeabilized using a Kryptonite permeabilization buffer comprising 6% Sarkosyl, 1M urea, and 4ug/ul proteinase K. The reverse transcriptase reaction was performed on human tonsil samples were prepared as disclosed above using six different conditions: (1) the 42B RT control variant was used with RT reagent; (2) the 42B RT control variant was used with Buffer X.5; (3) the engineered N-DAT-42B RT variant was used with the RT reagent; (4) the engineered N-DAT-42B RT variant was used with buffer X.4; (5) the engineered N-DAT-42B RT variant was used with buffer X.5; and (6) the small-scale purified engineered N-DAT-42B RT variant was used with buffer X.5. Most of the engineered N-DAT- 42B RT variants tested were purified by FPLC. However, some N-DAT-42B RT variants underwent a small scale purification. RT was performed using a 2 step protocol: 45 min at 53 °C and 30 min at 42 °C. [000580] As shown in FIGs.28A-F, the engineered N-DAT-42B RT variant maintained a good spatial resolution using all three buffers tested. FIGs.28A-F show UMI heat maps showing globally detected gene expression in a human tonsil tissue analyzed using N-DAT-42B RT variant (FIGs.28C-F) or 42B RT variant (FIGs.28A-B) and the three different buffer formulations. The NDAT1 RT variant was either FPLC-purified or underwent a small scale purification (FIG.28F). The global detection of gene expression is shown as a heat map. The RT Reagent (FIG.28A and FIG.28C) showed better mapping metrics when compared to buffer X.4 (FIG.28D), and buffer X.5 (FIG.28B, FIG.28E, and FIG.28F). The commercially available RT reagent comprised 2% Glycerol, 50 mM Tris pH 8.3, 3 mM MgCl2, 75.04 mM KCl, 0.5% Supersonic F-108, 0.489 mg/ml BSA, and 3.16 mM dNTPs. Buffer X.4 and buffer X.5 have the same components as the commercial RT reagent with the following differences. Buffer X.4 contains 1.05% Glycerol, 158.3 mM NaCl, 0.579 mg/ml BSA, 4.685 mM dNTPs, 1.5mM dCTP, and 0.775mM GTP. Buffer X.5 contains 133.8 mM NaCl, 1% Supersonic F-108, 3.96 mM dNTPs, 1.31 mM dCTP, and 0.655 mM GTP. A MMLV RT variant without a DAT fusion domain was used as a control (e.g., a RT variant comprising the amino acid sequence of SEQ ID 160 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    NO: 1, 143, 145, or 172). The NDAT1 RT variant was either FPLC-purified or underwent a small-scale purification (FIG.28F). In particular, FIGs.28A-F show that NDAT1 RT variant improved the sensitivity of the assay while maintaining good spatial resolution when compared to the control RT variant, with all three buffer conditions. [000581] FIGs.29A-F show that the engineered N-DAT-42B RT also increased the quality and the sensitivity metrics of the spatial assays. FIGs.29A-F show graphs comparing the quality, sensitivity, and detection of gene expression under the same six conditions shown in FIGs.28A- F. In particular, fraction reads in spots under tissue (FIG.29A) was substantially the same among the 6 conditions tested. However, the smallest standard deviation across the six conditions tested was obtained using RT reagent and with the engineered N-DAT-42B RT. Reads mapped confidently to transcriptome (FIG.29B) were significantly increased in the samples containing the N-DAT-42B using the RT reagent, as compared to using buffers X.4 and X.5, and compared to all three buffer conditions using the RT control variant. The total genes detected (FIG.29C) appeared the same in all conditions tested, with the exception of the RT control variant using the RT reagent (RTR), which showed significantly reduced total genes detected. Analysis of median UMI counts per spot (30k mapped spot-reads per spot, mapped to GRch38 reference genome assembly) (FIG.29D) showed that median UMI counts per spot were significantly enhanced in all three conditions using an N-DAT RT variant when compared to the three conditions using the RT control variant. Unexpectedly, N-DAT RT in bufferX.4 (N- DAT_RTX4) showed less variability. In addition, the GRch38 median genes per spot (30k mapped spot-reads per spot) (FIG.29E) analysis showed that median genes per spot were significantly enhanced in all samples comprising an N-DAT RT variant when compared to the 42B RT samples regardless of the buffer used. As such, engineered NDAT1 RT variant improved the sensitivity of the spatial assay while maintaining good spatial resolution when compared to the 42B RT variant (FIGs.29D-E). The RT Reagent appeared to show better mapping metrics when compared to Buffer X.4, and Buffer X.5. RT Reagent B with N-DAT at various concentrations and timings can also be tested and will likely show similar results as the RT Reagent. [000582] Assessment of individual gene expression also showed substantially similar results as globally detected gene expression. FIGs.30A-F show UMI heat maps shown as a log10(UMI) 161 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    ranging from cold to hot (0, 0.5, 1.0, 1.5, and 2.0), which illustrate the gene expression of RGS3 in a human tonsil tissue analyzed using N-DAT RT variant (FIGs.30C-F) or 42BRT control variant (FIGs.30A-B) under the three different buffer formulations (RT reagent (FIG.30A and FIG.30C), Buffer X.4 (FIG.30D), and Buffer X.5 (FIG.30B, FIG.30E, and FIG.30F)). [000583] FIGs.31A-C show UMI heat maps shown as a log10(UMI) ranging from cold to hot (0, 0.5, 1.0, 1.5, 2.0, and 2.5) illustrating the gene expression of KRT5 in a human tonsil tissue analyzed using N-DAT RT variant (FIGs.31C-F) or RT control variant (FIGs.31A-B) using three different buffer formulations (a commercially available RT reagent (FIG.31A and FIG. 31C), Buffer X.4 (FIG.31D), and Buffer X.5 (FIG.31B, FIG.31E, and FIG.31F)). Under sandwich configuration conditions used in the assay shown in FIGs.30 and 31, some evidence of undesirable flow of transcripts and/or target molecules or analytes was observed. A black circle on upper right quadrant identifies images with blurry tail. Example 11: Spatial resolution in non-human/mouse tissue. [000584] To further determine the effectiveness of the engineered N-Dat RT variant (e.g., 42BL-NDAT), a spatial assay was performed using a zebrafish sample using Visium standard definition (SD) spatial gene expression assay (e.g., with a first slide configuration) or Visium high definition (HD) spatial gene expression assay. Visium HD slides contain two 6.5 x 6.5 mm Capture Areas with a continuous lawn of oligonucleotides arrayed in ~11 million 2 x 2 µm barcoded squares without gaps, achieving single cell–scale spatial resolution. The data are output at 2 µm, as well as multiple bin sizes. The 8 x 8 µm bin is the recommended starting point for visualization and analysis. See e.g., 10xgenomics.com/platforms/visium; 10xgenomics.com/products/visium-hd-spatial-gene-expression. [000585] The zebrafish samples were permeabilized using a Kryptonite permeabilization buffer comprising 6% Sarkosyl, 1M urea, and 4ug/ul proteinase K. The reverse transcriptase reaction was performed on zebrafish samples using T reagent, the engineered N-DAT-42B RT (NDAT- 42BL) variant, TSO, and dTT. The RT was performed using a 2 step protocol-45 min at 53 °C and 30 min at 42 °C. The single strand synthesis was performed using a control RT (e.g., MMLV TR variant comprising an amino acid sequence of SEQ ID NO: 1, 143, 145, or 172), RT Reagent, and SS Primers. Under current experimental conditions, N-DAT fusion did not sufficiently process the synthesis of the second strand cDNA in the Zebrafish tissue. This was 162 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    likely caused by the N-DAT fusion configuration or the RT variant configuration. For example, the configuration may have slowed the enzyme from initiating or continuing the synthesis of the second strand. However, the N-DAT fusion sufficiently process the synthesis of the first strand cDNA in the Zebrafish tissue. [000586] FIGs.32A-B show UMI heat maps shown as a count log10 ranging from cold to hot (0.5, 1, 1.5, 2, 2.5, 3, 4, 4.5, and 5), illustrating globally detected gene expression in Zebrafish (non-human or mouse tissue) using a spatial assay with a first slide configuration. FIGs.32C-D show heat maps illustrating individual gene expression. FIG.32C shows Crgm1 gene expression as log normalized per experiment ranging from cold to hot (0.00-5.0). FIG.32D shows KRT5 gene expression as log normalized per experiment ranging from cold to hot (0.0-7.0). FIG.32E shows Rho gene expression as log normalized per experiment ranging from cold to hot (0.0-8.0). The quality, and sensitivity metrics were: fraction reads in spots under tissue (0.9), fraction reads usable (0.4), Reads mapped confidently to genome (0.56), and number of reads (250M). Fraction reads mapped to genome are the fraction of reads that mapped to a unique gene in the genome. In general, the read must be consistent with annotated splice junctions and are considered for UMI counting. [000587] FIGs.33A-D also show UMI heat maps showing globally detected gene expression (FIG.33A) or individual gene expression (FIGs.33B-C) in Zebrafish (non-human or mouse tissue) analyzed using NDAT1 RT variant (42BL-N-DAT1) and RT reagent buffer under sandwich configuration conditions for a spatial assay with a second slide configuration. FIG. 33B shows Crgm1 gene expression as log normalized per experiment ranging from cold to hot (0.00-4.0). FIG.33C shows KRT5 gene expression as log normalized per experiment ranging from cold to hot (0.0-9.0). FIG.33D shows Rho gene expression as log normalized per experiment ranging from cold to hot (0.0-7.0). The quality, and sensitivity metrics were: fraction reads mapped to genome (0.72). [000588] The engineered N-DAT RT variants disclosed herein enhanced the quality and the sensitivity of the spatial assay metrics while maintaining good spatial resolution when compared to a control RT variant. The control RT variant lacked the N-DAT but comprised the same RT backbone as the engineered N-DAT RT variant. The control RT variant consisted of the amino acid sequence of SEQ ID NO: 1, 142, 143, or 172. 163 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    EQUIVALENTS [000589] The present technology is not to be limited in terms of the particular embodiments described in this application, which are intended as single illustrations of individual aspects of the present technology. Many modifications and variations of this present technology can be made without departing from its spirit and scope, as will be apparent to those skilled in the art. Functionally equivalent methods and apparatuses within the scope of the present technology, in addition to those enumerated herein, will be apparent to those skilled in the art from the foregoing descriptions. Such modifications and variations are intended to fall within the scope of the present technology. It is to be understood that this present technology is not limited to particular methods, reagents, compounds compositions or biological systems, which can, of course, vary. It is also to be understood that the terminology used herein is for the purpose of describing particular embodiments only and is not intended to be limiting. [000590] In addition, where features or aspects of the disclosure are described in terms of Markush groups, those skilled in the art will recognize that the disclosure is also thereby described in terms of any individual member or subgroup of members of the Markush group. INCORPORATION BY REFERENCE [000591] All publications, patents, and patent applications mentioned in this specification are herein incorporated by reference to the same extent as if each individual publication, patent, patent application, or item of information was specifically and individually indicated to be incorporated by reference. To the extent publications, patents, patent applications, and items of information incorporated by reference contradict the disclosure contained in the specification, the specification is intended to supersede and/or take precedence over any such contradictory material. 164 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    [000592] Table 1 shows listing of non-limiting embodiments of RT enzymes of the present disclosure. [000593] Tables 2 and 5 shows additional listing of amino acid and nucleic acid sequences of non-limiting embodiments of the engineered RTs of the present disclosure. Table 1. Non-limiting embodiments of RT variants of the invention SEQ ID NO Description Sequences S L S P G V K H P K L T
Figure imgf000166_0001
165 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    Table 1. Non-limiting embodiments of RT variants of the invention SEQ ID NO Description Sequences I K E G E L I A E N A
Figure imgf000167_0001
166 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    Table 1. Non-limiting embodiments of RT variants of the invention SEQ ID NO Description Sequences aa ga tc tc cc g at c
Figure imgf000168_0001
167 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    Table 1. Non-limiting embodiments of RT variants of the invention SEQ ID NO Description Sequences g c g c I V P GI I A T T P K W Q Y H P I V P GI I A T P K W Q Y H I V P GI I
Figure imgf000169_0001
168 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    Table 1. Non-limiting embodiments of RT variants of the invention SEQ ID NO Description Sequences A T P K W Q Y H I P P G P F L P H S G P I P P G P F L P H S G P
Figure imgf000170_0001
169 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    Table 1. Non-limiting embodiments of RT variants of the invention SEQ ID NO Description Sequences I V P GI I A T P K T W Q Y H I V P GI I A T P K T W Q Y H V F S F K L
Figure imgf000171_0001
170 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    Table 1. Non-limiting embodiments of RT variants of the invention SEQ ID NO Description Sequences E V F S F K L E M V F S F K L E N A V F S
Figure imgf000172_0001
171 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    Table 1. Non-limiting embodiments of RT variants of the invention SEQ ID NO Description Sequences F K L E V F S D C Q G A G A V F S D C Q G A G K
Figure imgf000173_0001
172 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    Table 1. Non-limiting embodiments of RT variants of the invention SEQ ID NO Description Sequences D V F S F K L E I V P GI I A T P K T W Q Y H I V P GI I A T P
Figure imgf000174_0001
173 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    Table 1. Non-limiting embodiments of RT variants of the invention SEQ ID NO Description Sequences K T W Q Y H I P P G P F L P H S G P I P P G P F L P H S G P I V P GI I A
Figure imgf000175_0001
174 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    Table 1. Non-limiting embodiments of RT variants of the invention SEQ ID NO Description Sequences T P K W Q Y H I V P GI I A T P K T W Q Y H I V P GI I A T P K W Q Y H I V
Figure imgf000176_0001
175 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    Table 1. Non-limiting embodiments of RT variants of the invention SEQ ID NO Description Sequences P GI I A T P K W Q Y H I V P GI I A T P K W Q Y H I P P G P F L P H S G
Figure imgf000177_0001
176 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    Table 1. Non-limiting embodiments of RT variants of the invention SEQ ID NO Description Sequences P I P P G P F L P H S G P I V P GI I A T P K W Q Y G I V P G I A T P K
Figure imgf000178_0001
177 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    Table 1. Non-limiting embodiments of RT variants of the invention SEQ ID NO Description Sequences W Q Y G I V P GI I A T P K W Q Y H I V P G I A T P K W Q Y H I V P GI I A T
Figure imgf000179_0001
178 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    Table 1. Non-limiting embodiments of RT variants of the invention SEQ ID NO Description Sequences P K W Q Y H L L R L Q E F R L A K Q L L R L Q E F R L A K Q
Figure imgf000180_0001
179 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    Table 1. Non-limiting embodiments of RT variants of the invention SEQ ID NO Description Sequences I P P G P F L P H S G P S P P G P F L P H S G P K K P IS K P P
Figure imgf000181_0001
180 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    Table 1. Non-limiting embodiments of RT variants of the invention SEQ ID NO Description Sequences T W Q Y H P
Figure imgf000182_0001
181 4876-6828-0003.1 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000183_0001
u n S d G QV EQE F T N LP K MS LTQ YG C GP P V d i c TY n A EQCRH KP L S AP F S VGK NY LQYL AAV P T LLAGLGVS HA TYS QE QL LR P LQYDKWVHLAP N a o AI L T I I R R EQCR n i WS VDP VL I QP L P P KL TYTPKG QLGR VA Q F L I QGKA E S S LKLS LA AKI P L AI E KKAGPAGAKWRANI WS VD di m APGVG S GL L LLP a T QEKDA P P G F L I NKGGKQL AVL I Q c A QTQP RS TQP DAYP M EYDS LKQQDTAER RS HR QP TGVP o F S n i DT DYL S AL L KLDLP LP RKQ RNNH Y L RH AAVYT LKVLYW I ARL FAS P F S QR RT Q TV I GQGFKF KMG LAVV RC WVMYI L MN P DT D PQE VA LY I L DNT HE K EAR S S LA KLD me a r W L P Q I TP LWR FQQVI QF P PAA ETHDT EL N AGK S LR N T f T I I HG P N P C F T L DQ CC I T E F L D P KR R A VD L Q L A D QH I L L G RN E W T P I Q I G o u s g o l ni c t si n oi s t 6 1 si 1 0 H 0 0 i d L t p i r x T D -g x a T D 1 . : n 2 e c s R L s e O T - R L O 3 0 0 el e D r S N S 0-8 b p 2 8 a e Q : 6 T h t E S D I O 3 N 6 5 6 -6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000184_0001
u S WTD EL LKAMS N T TY P PN n d i W QE F PVNKQ P LAVT LLGQ LGVK S C HA GTYS V d c HA S NYGLL LRY P LAYP L A KWVHLANT I I K EQQE n A a o S d n F i P KQTYTPQ KGD i m P L LGR P P F L LQGKAI QKAGVAS P P ES LL KLS R A AKC P R L QGI L L LPAE K EKDAGALGA WLAN WI S VD NKK GG R KQI L AVL I Q ca A G S S QD L AYT P MQ EYDS KP QF I TAER RPDR QP TGV o LTP n i LP LP RKQ NHRH LT Q TV AAV L YT QD KVLYWARLAS P F S QP R I GQGFKF KMG LAL VV RC W I VMYF I L MN P DT DY PQE VA I L DNT EK EAR S S LA KL LD me a r Y P RWR FQQVI QF P PAA ETHH DTL N AGK S LR S T f NL CT L DQ CC I T E F L D P KR R A VD L Q L AE T QH I L L G RN E W T P Q o u s I I G g o l ni c t si n oi s s t . i 2 H 0 i d x 0 L t p i nr c GI : s F g a T t R D L 1 . 3 : 2 e s e - 0 N O 0 el e D r S 0-8 b p 2 8 a e Q : 6 T h t E S D I O 7 N 6 -6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000185_0001
u n S d i W QE T F P S VN LP GK LQYLK MS LTQTYK C G QVQE AAV P T LLAGLGVS HA TY QC EHF T AP S V d c A HANY QL LR P LQYDKWVHLAP N L T I I K R EKP R L S FNY n a o S d n i P F i P L P K F L T L YTPKG QLGR VA I QGKAI E S S LKLS E KKAGPAGAKWLA AI VDP RAN WS VL I QP L LP KQ F L L c m a A L GQG S S QL L P D LLPA QEKDA AYT P M EYDS LKP GLNKGGKQI L APGVGQ S GI L QF QI DTAER RPDR QTQP S TQD o LT n i LP LP RKQ NHRH LT Q TV AAVYTKVLYWARL FAS P F S T D R L LP P LP I GQGFKF KMG LAL VV RC W I VMYL MN P DALY DNH L RH PQE VA L DNT EKI EAR S S L KL RS YR T Q I me Y P RWR FQQVI QF P P I AA ETHH DTL N AGK S L Q I TP LWR a r NL CT L DQ CC I T E F L D P KR R A VDQ E HLGN WPI GNCT F f u L L A T Q I L R E T I H P P F L D o s g o l ni c si n oi si h 3 0 si tsi d t L t p i : g x 0 H : nr T D : g 1 . 2 e c s a s e T - R L a D N O t- 3 0 S N 0 0 e - l e r 8 b p 2 a e Q : 8 6 T h t E S D I O 9 - N 6 6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000186_0001
u n S d N LP LK MS LTQTYK C G QVQE F T LP d i c GK n A LLQ LRY P LAAV QYP T D LLAGLGVS HA TY KWVHLANT I I K R EQC E P R LHS AP F S VN GKQY NYLL LR P L a o n TYTP i QGKAKG I QLGR VA E S S P LL KLS A AKI VDP LKQ E L TYTP KKAGPAGAKWL RAN WS L I QP LP F LQGKA di m L LLPA QEKDAGLNKGGKQI L AVPGVGQ S GI L L LPA ca A AYT P M EYDS KP QF I TAER RPDR QTQP S TQD L AYT P M o RK n TVQAAV L YT QD KVLYWARLAS P F S D R L P P LP RKQE A i GQGFKFAL MGQL E VVV A LR DC NW I F TVMY EKI L MN DT EARP S S ALYL KL RDS NH Y L RHTVGF RT Q I GQMG me QK a r QQVP I QF CC I T E F L D P P P I AA THHTLGKNS LL Q TP LWRQK QVP I KR E DQDEAHLGN WPI I GNCT F QCTF f u s R A V L L A T Q I L R E T I H P P F L D C I E L o g o l ni c t si n oi 4 si si d t 0 pi x 0 h T D - L t : nr g 1 2 e c s s e R L a . O t- 3 0 0 el e D r S N 0-8 b p 2 a e Q : 8 6 T h t E S D I O 1 - N 7 6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000187_0001
u n S d LK MS LTQTYK C G QVQE F T N LP K M d i c AAV n A QYP T D LLAGLGVS HA TY KWVHLANT I I K R EQC E P R LH P S A F S VGKQYL AAV NYLL LR P LQYP T D L K a o d n i KG I QLGR VA i KKAGP E S S P LL KLS A AKI VDP LKQL TYTPKGLG AGAKWL RAN WS L I QP LP F LQGKAI Q KKA c m E a A QEKDAGLNKGGKQI L AVPGVGQ S GI L L LPAEEKD YDS LKP QF I TAER RPDR QTQP S TQD L AYT P MQ YDS K o AV n i KFYT QD K QLAL VLYWARLAS P F S D R L P P LP RKQE AAV L YT EVVV LR DC NW I TVMYF EKI L MN DT EARP S S ALYL KL RDS NH Y L RHTVGFKFAL RT Q I GQMGQL E VV me QF P P A I AA THHTLGKNS LL Q TP LWRQK QVP I QF P A I a r D P KR R AE VD L QDEAHLGN WPI I GNCT F QCTF D P RA f u s L A T Q I L R E T I H P P F L D C I E L P K R A o g o l ni c t si n oi 5 si si d t 0 pi x 0 H T D : g 1 L t : nr 2 e c s s e R L a . O T- 3 0 0 el e D r S N 0-8 b p 2 a e Q : 8 6 T h t E S D I O 3 - N 7 6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000188_0001
u n S d S LTQTYK C G QV EQE T LP K MS LTQ d i c LAGLGVS HA TY n A WVHLANT I I K R EQCRHF P S VN GKQYL AAV KP L S A FNYLL LR P LQYP T D LLAGL KWVHL a o d n R i GVA P E S S P LL KLS A AI VDP LKQL TYTPKGLGR VAS S i AGAKWL RAN WS L I QP P P F LQGKAI Q KKAGP E AG c m AG a A P LNKGGKQI L AVPGVGQ S GI L L LPAEEKDAGLN QF QI DTAER RPDR QTQP S TQD L AYT P MQ YDS KP QF I T o Y L S P S R L P P K E L T QD K ni VV LR L D C NWW I A TVMR A YF EKI L MN F D DT Y LP R Q AVY K LY EARP S S ALDL KLNNH Y L RHTVGA FKFAL VV R T Q I GQMGQL RC E W VVL DNT me A ETHH DTLGKNS LLR QTP LWRQK QVP I QF P A I A THH a r VD L Q L AE TA QHLGN WPI I GNCT F QCTF D P RAE DQD f u s I L R E T I H P P F L D C I E L P K R A V L L A o g o l ni c t si n oi 6 si si d t 0 pi x 0 H T D : L t : nr g 1 2 e c s s e R L a . O t- 3 0 0 el e D r S N 0-8 b p 2 a e Q : 8 6 T h t E S D I O 5 - N 7 6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000189_0001
u n S d TYK C G QV EQE T LP K MS LTQTYK C d i c GVS HA TY n A ANT I I K R EQCRHF AP S VN YGKQYL AAV KP L S FN LL LR P LQYP T D LLAGLGVS H KWVHLANT I I a o P d n i LL AKLS A AI VDP LKQL TYTPKGLGR VA i KWL RAN WS L I QP LP F LQGKAI QKAGP E S S P LL KLS AGAKWL R c m K a A AGG W ERKQI ARPDL AV RLAR QP S P TGV QP GQ S S GI F S D R TQL L LPAE K D L EKDAGLNKGGK AYT P MQ YDS KP QF I TAER YL P P LP RKQEAV L YT QD KVLYWARP RL o n I M i V YF EK L MN DT ALDL NH L RHTVGA FKFAL VV RC W I MYF L L I EARP NS S S L KL RN T Y P R T Q I GQ QKMGPQL QE V P A I L AD TNTV HHEK L I EA me TAGK L Q I LWR QVI F T GK a r E T QH I L L G RN E WPI GNCT F QCTF D P RAE DQDEAHL f u s T I H P P F L D C I E L P K R A V L L A T Q I L o g o l ni c t si n oi 7 si si d t 0 pi x 0 H T D -g 1 L t : nr 2 e c s s e R L a O t . _ 3 0 0 el e D r S N 0-8 b p 2 a e Q : 8 6 T h t E S D I O 7 - N 7 6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000190_0001
u n S d i A GTYQV E T E RLLG KNR E S TGC QC EQ HF AP S VN GKLP LKAMS LTQTYK S HA d c n A K R EKP R L S FNYLLQ RYAAV P T LLAGLGVT I I K R a o d n i A AI VDP L AN WS KQL TL P L PQY GDKWVH ALAP N LLS i VL I QP P P F LQYTAKI QLGR V E S S LKWLA RAN c m QI a A DL A R QP TGV QP GQ S S GI TQL LG D LL KAE KKAGPAGAKGKQI L AYP TMQEKDA P G F L I NKG R PDR o AS P n i M RN F S T D R L LP P LP P D RKP EYDS LKQQDTAE R L S S AL FAS L KLY RDS NH Y L RHTVQAAVYT R T Q I GQGF LKVLYW I ARL MN P GKF LAVV RC WVMYI AR S me NS L Q I TP LWR FQKMP I Q QEVA I L D TN HTE K EKNS a r G RN E W T PI I HG P N P C F T L DR C QV F P A HT LGLGN f o u s C T H D P R A E D Q D E A H L R E g o l ni c si n oi 8 0 si tsi d t L t p i x 0 H T D : g 1 . : nr 2 e c s s e R L a O t- 3 0 N 0 el e D r S 0-8 b p 2 8 a e Q : 6 T h t E S D I O 9 N 7 -6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000191_0001
l c q V u e S GM P P S NY F VWW DG EL T TDL G E R L L T LGAS AP G KAMK R TKGT I S N TE S TYGP CA n d i GTYQ CE RQF P S VNKQ P YLAVT LLL Q AGLGVG S HA d c n A EQP LH a o AKI S A FNYGLL LR P LA QYP KWVHLAP NT I I R R S VDP P LKQL TYTPKGD QLGRVA E S S LL KLS LA d n i W ic m AVL P I Q GVP P QF L I Q a A QP T P G S G LGKA PAI E KKAEPAGAKWRANI S QRS LT PQL P D LL YT PAKP MQEK EYDS DA LKP G F L I NKGG RKQL TQQDTA Y E RP LHR o F T D n LYL H L i DS ALDNL RHR TQ TVQ GA FAV FYLKVL C WA I GQ K W I MR FAS N MGQLAVV LR DNT V E KYI L M ARP me L K a r W LR N T Y P R LWRQKVP I QE F V P A I A THHT LE GKNS S T PI Q I I HG P N P CT F QQ CTF D P RAE DQDEAHLGN f F L D C I E L P K R A V L L A D Q I o u s L R E g o l ni c t si n oi 9 0 si si d t L t p i x 0 H T D : g 1 : nr 2 e c s s e R L a . O T- 3 0 0 el e D r S N 0-8 b p 2 a e Q : 8 6 T h t E S D I O N 1 8 - 6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000192_0001
u S WYREH S G QYAAP LL HLANT I I R G QV n d d i c TQC n A EKI P R VL S A DP F LNYLL KQL TLR YP T L Y PQ DKWVAS P KGLGR V P E LLLS R TY CE R AS AKWLA A EQ KP L a o A d n S i WVL i m AP I QP P P F LQGKAI TGVGQGI QP S S TQL L LPAEQKAGGLGKKG R KQNI AI D L QK EK AYT P MYDS DA F I N LKP TAG ER S HL WS VLD I Q QQDYWAR RLAR AVPGV ca A QP S D R L LP P LP RKQE AAVYTKVL CW I MYF L M S QTQP R o F T ni DALYNHRHTVGFKFAL VV RNT V S L KLD E LR S YL Q TP RT Q LWI GQ KI R KMGQL E VA L D THHT LEAR N P P F S GKNS DT DY AL LD FQQVP I QF P RI ADQDEAHLG S N S L K R N me P a r W I I G TI L H P P N KP C T VF F LDQ Q AC C DI T QEF D KL RP P RA AELLADQI LRE L Q T GK T V W LGD T THAAI WPI I G f u s V A L E P V L A L E P T I H P o g o l ni c t si n oi s d t 0 1 s pi x 0 i 1 H 1 : x 0 i L t r T D g T D 1 . : n 2 e c s s e R L a O t- R L 3 0 N O 0 el e D r S S 0-8 b p 2 8 a e Q : 6 T h t E S D I O 3 N 8 5 8 - 6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000193_0001
u n S d i QE HF T AP S VN LP GK LQYLK MS LTQ YGC G QV AAV P T LLAGLGVKHA TY CE RQE T HF AP S V d c A S FNYL LR P LQYDKWVHLAP N L S TI I K R EQ KP L S FNY n a o d n P L i P i P P KQ F L T L YTPKG QLGRVA I QGKAI E S S LKL S E KKAEPAGAKF LA AI RAN WS VDP LKQL VL I QP P P F L I c m a A GQ S S G TQL L D LLPA QEKDA AYT P M EYDS LKP GLDKGGKQI L APGVGQ S GL QF QI DTAER RPDR QTQP RS TQD o L LP P LP RKQ ni NH Y L RHTV AAVYTKVLYW I ARLAP P F S T DYL LP P LP RT Q I GQGFKF KMG LAL VV RC WVMYF L MN P DALDNH L RH PQE VA L DNT GKI EAR S S L KL RNYR T Q I me a r P NLWR P C F T LFQQVI QF DQ CC I T E F L D P P P I AA KR R AETHH DTL N AGK S L Q VDQ E QLGN WPI I T GP NL CW TR F f u L L A T Q I L R E T I H P P F L D o s g o l ni c si n oi si 2 H 1 si tsi d t L t p i : g x 0 H : nr T D : g 1 . 2 e c s a s e T D - R L a N O t- 3 0 S N 0 0 e - l e r 8 b p 2 a e Q : 8 6 T h t E S D I O N 7 8 - 6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000194_0001
u n S d N LP LK MS LTQ YG C WYREH S GLQY d i c GKQYAAV n A LL LR P LQYP T LLAGLGVS HA TQC KWVHLAP NT I I R R EKI P R VL S ANYL LR DP F LKQTYP T L P a o n TYTP i QGKAKGD I QLGR VA KKAGP E S S LL KLS A AS LQP P P F L LQGKA AGAKWL RAN WVP I GVGQGI L L LPA di m L LPAE QEKDAGLNKGGKQI L ATQP S S TQD LYT P M c A L AYT P MYDS KP QF I TAER RS HR QS D R L P P LPA RKQE a o RK n TVQE GAAV L FYT QD LKVLY C W I ARL FAS P N F T ALYL LDNH LRH TQ TVGA GQ F i GQ FK AVV R WVMYL M D S Y I MG QKMGPQL E V P A I L AD TNT HHE TK L I EARP NS S S L K LR QTP R LW TR FQK QQVP I me a r QQVI QF CC I T E F L D P P GK P I I GN C CT F KRAE DQDEAHLGN WI H P P F LDCI E L f u s R A V L L A D Q I L R E T L P K V F Q A D Q K R o g o l ni c t si n oi 3 si si d t 1 pi x 0 H T D : L t : nr g 1 2 e c s s e R L a . O t- 3 0 0 el e D r S N 0-8 b p 2 a e Q : 8 6 T h t E S D I O N 9 8 - 6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000195_0001
u n S d AAP LL GLGVS H G QV EQE F T LP K M d i c QYDKWAHLANT I I R n A KGLGRVAS a o I QK S P LLLS R TY A EQC P R LH P S A F S VN GKQYL AAVT NYLL LR P LQYP DL K d n i E AE V P E AGAKWL RA AKI VDP LKQL TYTPKG QLG i QK EKDAGLNKKGKQNI WS L I QP P P F LQGKAI KKA c m a A YDS LKP F I TAG ER PHL AVPGVGQ S GI L L LPAEEKD AVYTQ KQDLYWAR RLAR QTQP S TQD L AYT P MQ YDS K o KFAL ni QLVVVV LRC W I MYF QE F P A I ADNT V L THHE M S TKI LEAR N P P F S D R DT YL P P LP RKQEAV L YT ALDL NH GKNS S KL RNY L RHTVGA FKFAL RT Q I GQMGQL E VV me D P P R RAEDQDEAHLG S N LL Q TP LWRQK QVP I QF P A I a r GK T AV W V AL LLADQ I LR E WPI I GNCT F QCTF D P RA f u s G D T T H A A I T I H P P F L D C I E L P K R A o g o l ni c t si n oi 4 si si d t 1 pi x 0 H T D : g 1 L t : nr 2 e c s s e R L a . O T- 3 0 0 el e D r S N 0-8 b p 2 a e Q : 8 6 T h t E S D I O N 1 - 9 6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000196_0001
u n S d S LTQ YG C G QV EQE F T LP K MS LTQ d i c LAGLGVS HA TY n A WVHLAP NT I I R R EQCRH KP L S AP F S VN GKQYL AAV NYLL LR P LQYP T D LLAGL KWVHL a o d n i R E VA P E S S LL KLS A AI VDP LKQL TYTPKGLGRVA i AGAKWL RAN WS E S S VL I QP P P F LQGKAI Q KKAEPAG c m AG a A P LNKGGKQI L APGVGQ S GI L L LPAE QEKDAGLD QF QI DTAER RPHR QTQP S TQD L AYT P MYDS KP QF I T o Y L S P S R L P P P K E L T QD K ni VV LR L D C WA NW I T VMR A EKYF I L MN F D DT Y L R Q AVY K LY EARP S S ALDL KL RNNH Y L RHTVGA FKFAL VV R T Q I GQMGQL RC E W VVL DNT me A ETHH DT ELGKNS LL Q TP LWRQK QVP I QF P A I A THH a r VD L Q L A A D QH I LGN WPI I GNCT F QCTF D P RAE DQD f L R E T I H P P F L D C I E L P K R A V L L A o u s g o l ni c t si n oi 5 1 si si d t p i x 0 H T D : g 1 L t : nr 2 e c s s e R L a . O T- 3 0 0 el e D r S N 0-8 b p 2 a e Q : 8 6 T h t E S D I O N 3 9 - 6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000197_0001
u n S d YGC WYREH S G QYAAP LL GLGVKH d i c GVKHA TQC A AP NS TI I K R EKI P R VL S A DP F LNYLL KQTLR YP T L Y PQGD LKWAHLANS GRVAS P LLTI I L S n o LL KL S A AS LQP P P F L LQGKAKI QKAE V P ES GAKF L a d n i A i KF L RAN WVI GVGQGI L L LPAE KKDAGALDKKG R K c m K a A AGG ERKPQI DL AP R QT S QP S S Q DR LT P P L D LYT PAKP MQE DS QEY LKP F I TAG ER P AVYTQQDLYWAR RL o W I AR RLAP P F TLYLHRHR TVGA FKFALK VV C W I MYF L ni VMYF GK L MN DALDS N YL T Q I GQMGQLVVLR DNTVKI EA LI EARP NS S S L K LR QTP R LW TR FQK QQV TP I QE F P A I P RAA ETHHG QDT L AGK me T GK P I I GN C C F D R D E QL a r E TA QQ I L L G RN E WI H P P F LDCI E L P K AVL LATQ I L f u s T L P K V F Q A D Q K R G T W V A L G D T T H A o g o l ni c t si n oi 7 si si d t 1 pi x 0 H T D : g 1 L t : nr 2 e c s s e R L a . O T- 3 0 0 el e D r S N 0-8 b p 2 a e Q : 8 6 T h t E S D I O 5 - N 9 6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000198_0001
u n S d i K G R TYQV C E RQE HF T L AP P K M S VN S LTQ YG C GKQYL AAV P T LLAGLGVS HA d c n A A EQ KP L S FNYLL LR P LQYDKWVHLAP N a o L T I I R R d n i A AI QN i I WS VDP VL I QP L P P KQ F L T L YTPKG QLGRVA I QGKA E S S LKLS LA AI E KKAEPAGAKWRAN c m DL A a A AR PGVGQ S GL L P QTQP RS TQD LLP AYT P MQEKDA EYDS LKP GLNKGGKQI L QF QI TAER RS HR o M n R N P P F S i NS DT DYL AL S S L KLDLP P LP RKQ RNNH Y L RHTV AAVYT D KVLYW I ARLAS RT Q I GQGFKF KMG LAL VV RC WVMYF L MN P PQE VA L DNTE KI EAR S me G RN L Q I TP LWR FQQVI QF P P I AA ETHH DT LGKNS a r AE I W T PI I HG P N P C F T L DQ CCTF D R DQ EAHLGN f o u s I E L P K R A V L L A D Q I L R E g o l ni c t si n oi 8 1 si s H i d t L t p i x 0 T D : g 1 . : nr 2 e c s R L a T 3 s e O - 0 0 el e D r S N 0-8 b p 2 8 a e Q : 6 T h t E S D I O 7 N 9 -6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000199_0001
l c q V u e S GM P P Y F S N DG VWW EL TDL G E R T L L T LGAS AP G KAMKS NR TE S TTKGT I YGP CA n d i GTYR CEQ HF P S VNKQ P YLAVT LLL Q AGLGVG S HA d c n A EQP R a o AKI L S A FNYGLL LR P LA QYP KWVHLAP NT I I R R S VDP P LKQL TYTPKGD QLGR VA E S S LL KLS LA d n i W ic m AVL I QP P QF L I QLGKA PAI E KKAGPAGAKWRANI a A QPGVG S G P T S QP RS LT PQL P D LL YT PAKP MQEK EYDS DA LKP G F L I NKGG RKQL TQQDTA Y E RS LHR o F T D n LYL H L i DS ALDNL RHR TQ TVQ GA FAV FYLKVL C WA I GQ K W I MR FAS N MGQLAVV LR DNT V E KYI L M ARP me L K a r W LR S T Y P R LWRQKVP I QE F V P A I A THHT LE GKNS S T PI Q I I HG P N P CT F QQ CTF D P RAE DQDEAHLGN f F L D C I E L P K R A V L L A D Q I o u s L R E g o l ni c t si n o 9 1 si si di t L t p i r x 0 H c T D : g 1 : n 2 e s s e R L a O t . - 3 0 0 el e D r S N 0-8 b p 2 a e Q : 8 6 T h t E S D I O N 9 9 - 6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000200_0001
u S WYREH S G QYAAP LL HLANT I I R G QV n d d i c TQC n A EKI P R VL S A DP F LNYLL KQL TLR YP T L PQYDKWVAS P KGLGR V P E LLLS R TY CE R AS GAKWLA A EQ KP L a o A d n S i WVL i m AP I QP P P F L TGVGQGI Q QP S S TQL LG L K PA AI EQKAGGL KKG R KQNI AI QK a P D L EK AYT P MYDS DA F I N LKP TAG ER PHL WS VD VL I Q QQDYWAR RLAR APGV c A QP S D R L LP LP RKQE AAVYTKVL CW I MYF L M S QP TQP R o F T ni DALY SL KLDNHRHTVGFKFAL VV RNT V E LR S YL Q TP RT Q LWI GQ KI R KMGQL E V EAR N P F S T DY P A L D THHT LGKNS DAL LD FQQVP I QF R I AA E DQDEAHLG S N S L K R N me P a r W I I G TI L H P P N KP C T VF F LDQ Q AC C DI T QEF D KL RP P RAVLLAD GK T W V ALGD TQ TI HL AR AE I W L PI Q I T G f L E P V L A L E P o u s T I H P g o l ni c t si n oi s t 0 2 si 1 i d x 0 H 2 0 L t p i r c T D : g x a T D 1 . : n 2 e s R L T R L 3 s e D O 0 S - N O S 0 0 e l e r - 8 b p 2 a e Q : 1 3 8 6 T h t E S D I O N 0 1 0 1 -6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000201_0001
u n S d i QE HF T AP S VN LP GK LQYLK MS LTQ YG C G QV AAV P T LLAGLGVS HA TY CE RQE T HF AP S V d c A S FNY QL LR P LQYDKWVHLAP N L T I I R R EQ KP L S FNY n a o d n P i P L i P P K F L T L YTPKG QLGRVA I QGKA E S S LKLS LA AI S VDP LKQL AI E KKAEPAGAKWRAN WVL I QP P P F L I c m a A GQ S S G TQL L P D LLP AYT P MQEKDA EYDS LKP GLNKGGKQI L APGVGQ S GL QF QI DTAER RS HR QTQP RS TQD o L LP LP RKQ ni NH Y L RHTV AAVYTKVLYW I ARL FAS P F S T DYL LP P LP RT Q I GQGFKF KMG LAL VV RC WVMYL MN P DAL LDNH L RH PQE VA I L DNTE KI EAR S S L K R NYR T Q I me a r P NLWR P C F T LFQQVI QF DQ CC I T E F L D P P PAA KR R AETHH DT VD L Q EL N AGK S HLGN W L PI Q I T GP NL CW TR F f u s L A D Q I L R E T I H P P F L D o g o l ni c si n oi s t i 2 s H 2 i tsi d L t p i r : g x 0 H T D : g 1 . : n 2 e c s s e at- R L a O t- 3 0 0 el e D N r S N 0-8 b p 2 a e Q : 8 6 T h t E S D I O 5 N 0 1 -6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000202_0001
u n S d N LP LK MS LTQ YGC G QV EQE F T LP d i c GK n A LLQ LRY P LAAV QYP T LLAGLGVKHA TY KWVHLAP NS TI I K R EQC P R LHS AP F S VN GKQY NYLL LR P L a o n TYTP i QGKAKGD I QLGR VA E S S LL KL S A AKI VDP LKQ E L TYTP KKAGPAGAKF L RAN WS L I QP P P F LQGKA di m L LLPA QEKDAGLDKGGKQI L AVPGVGQ S GI L L LPA ca A AYT RKP M QEYDS KP QF I TAER RPDR QTQP S TQD L AYT P M AAV L YT QD KVLYWARLAP P F S D R YL P P LP RKQE o n TVGF i GQ KFAL VV RC W I VMYF L MN DT ALDL NH KMGQL E VA L DNT GKI EARP S S KL RNY L RHTVGA F RT Q I GQMG me Q a r QQVP I QF CC I T E F L D P P P I AA KR R ET DHH QDT EL AGKNS L QLGN W L PI Q I T GP NL CW TR FQK QQVP CT I F f A V L L A T Q I L R E T I H P P F L D C I E L o u s g o l ni c t si n oi 3 2 si si d t L t p i x 0 H T D : g 1 : nr 2 e c s s e R L a . O T- 3 0 0 el e D r S N 0-8 b p 2 a e Q : 7 8 6 T h t E S D I O N 0 1 -6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000203_0001
u n S d LK MS LTQTYK C G R VQE F T N LP K M d i c AAV n A QYP T D LLAGLGVS HA TY KWVHLANT I I K R EQC E P R LH P S A F S VGKQYL AAV NYLL LR P LQYP T D L K a o d n i KG I QLGR VA i KKAGP E S S P LL KLS A AKI VDP LKQL TYTPKGLG AGAKWL RAN WS L I QP P P F LQGKAI Q KKA c m E a A QEKDAGLNKGGKQI L AVPGVGQ S GI L L LPAEEKD YDS LKP QF I TAER RPVR QTQP S TQD L AYT P MQ YDS K o AV n i KFYT QD K QLAL VLYWARLAP P F S D R L P P LP RKQE AAV L YT EVVV A LR DC NW I TVMYF EKI L MN DT EARP S S ALYL KL RDS NH Y L RHTVGFKFAL RT Q I GQMGQL E VV me QF P P I AA ETHHTLGKNS LL Q TP LWRQK QVP I QF P A I a r D P KR R A VD L Q LDEAHLGN WPI I GNCT F QCTF D P RA f A T Q I L R E T I H P P F L D C I E L P K R A o u s g o l ni c t si n oi 4 2 si si d t L t p i r x 0 H c T D : g 1 : n 2 e s s e R L a O t . - 3 0 0 el e D r S N 0-8 b p 2 a e Q : 9 8 6 T h t E S D I O N 0 - 1 6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000204_0001
u n S d S LTQTYK C G QV EQE T LP K MS LTQ d i c LAGLGVS HA TY n A WVHLANT I I K R EQCRHF P S VN GKQYL AAV KP L S A FNYLL LR P LQYP T D LLAGL KWVHL a o d n R i GVA P E S S P LL KLS A AI VDP LKQL TYTPKGLGR VAS S i AGAKWL RAN WS L I QP P P F LQGKAI Q KKAGP E AG c m AG a A P LNKGGKQI L AVPGVGQ S GI L L LPAE QEKDAGLN QF QI DTAER RPDR QTQP S TQD L AYT P MYDS KP QF I T o Y L S P S R L P P P K E L T QD K ni VV LR L D C NWW I A TVMR A YF EKI L MN F D DT Y L R Q AVY K LY EARP S S ALDL KL RNNH Y L RHTVGA FKFAL V R T Q I GQMGQL E VVV LR DC NW T me A ETHH DTL N AGK S LL Q TP LWRQK QVP I QF P A I A THH a r VD L Q L AE T QH I LGN WPI I GNCT F QCTF D P RAE DQD f L R E T I H P P F L D C I E L P K R A V L L A o u s g o l ni c t si n oi 5 2 si si d t L t p i r x 0 H c T D : g 1 : n 2 e s s e R L a O t . - 3 0 0 el e D r S N 0-8 b p 2 a e Q : 1 8 6 T h t E S D I O N 1 1 -6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000205_0001
u n S d YG C G QV EQE F T N LP K MS LTQ YGC d i c GVS HA TY Q n A AP NT I I R R E C P R LHS AP F S VGK NYLLQ LRYL P LAAV QYP T D LLAGLGVKH KWVHLAP NS TI I a o d n LL i AKLS A AKI VDP LKQL TYTPKGLGR VA i KWL RAN WS L I QP P P F LQGKAI Q KKAGP E S S LL KL S AGAKF L R c m K a A AGG W ERKQI ARPHL AV RLAR QP S P TGV QP GQ S GI L L LPAE QEKDAGLDK F S D RS TQD L GGK AYT P MYDS KP QF I TAER YL P P LP RKQE RP AAV L YT QD KVLYWARL o I ni VM EKYF I L MN DT EARP S S ALDL KL RNNH Y L RHTVGFKFAL V R T Q I GQMGQL RC E W I VVL DNTVMYF GKI L EA me T EL N AGK S LL Q I TP LWRQK QVP I QF P A I A THHTLGK a r D QH I L L G RN E W T PI GNCT F QCTF D P RAE DQDEAQL f I H P P F L D C I E L P K R A V L L A T Q I L o u s g o l ni c t si n oi 6 2 si si d t L t p i x 0 H T D : g 1 : nr 2 e c s s e R L a . O T- 3 0 0 el e D r S N 0-8 b p 2 8 a e Q : 6 T h t E S D I O 3 N 1 - 1 6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000206_0001
u n S d i A GTYR V C EQE HF T L AP P K MS S VN LTQTYK C GK LQYL AAV P T LLAGLGVS HA d c n A K R EQ KP R L S FNYL LR P LQYDKWVHLAP NT I I K R a o d n i A AI AN WS VDP LKQL TYTPKG QLGR VA E S S LL KLS i VL I QP P P F L I QGKA LA AI E KKAGPAGAKWRAN c m QI a A DL A R QP TGV QP GQ S S G TQL L D LLP QEKDA AYT P M EYDS LKP GLNKGGKQI L QF QI DTAER RPDR o AP P n i M RN F S T D R L LP P LP RKQ P D S S AL AAVYT LKLY RDS NH Y L RHTV RT Q I GQGFKF KMG LALKVLYWARL VV RC W I FAS VMYL MN PQE VA L DNT EKI EARP S me NS L Q I TP LWR FQQVI QF P P I AA ETHH DTL N AGK S a r G RN E W T PI I HG P N P C F T L DQ CCTF D R DQ E HLGN f o u s I E L P K R A V L L A T Q I L R E g o l ni c si n oi 7 2 si tsi d t L t p i nr x 0 H c T D : s R L g a 1 t . 3 : 2 e s e D O - 0 S N 0 0 e l e r - 8 b p 2 8 a e Q : 5 6 T h t E S D I O N 1 1 -6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000207_0001
l c q V u e S GMPNY FDGG TTGAGAP G TKGI GP S VWW EL TDLE R L LLKAMKS NR TE S TYGP CA n d i TYQ CE RQ HF P S VNKQ P YLAVT LLL Q K AGLGVS HA d c n A EQ a o AKI P L S A FNYGLL LR P LA QYP S VDP DKWVHLANT I I K R P L Q d n i W P KL TYTPKG QLGR VA E S S P LL KLS LA ic m AVL I QP Q F L I QLGKA PAI E KKAGPAGAKW GRANI a A QPGV P T P G S G QLLL YTMQEKS DA P G F L I NKG RKQL o F S QRS T D n LYLT LP P D H LPAKP EYDLKTQQDTA Y E RP LDR S i DS ALD RHR TVQ GA NNLT Q FAV FYLKVL C I GQ K WW I A MR FA MN MGQLA VVV LR DNTVKYI L ARP me L K a r W LR QT Y P R LWRQKVP I QE F P A I A THHE TLE GKNS S T PI I I HG P N P C F T F QQ CTF D P RAE DQDEAHLGN f u L D C I E L P K R A V L L A T Q I L R E o s g o l ni c t si n oi 8 2 si si d t L t p i x 0 H T D : g 1 . : nr 2 e c s s e R L a O t- 3 0 N 0 el e D r S 0-8 b p 2 a e Q : 7 8 6 T h t E S D I O N 1 1 -6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000208_0001
u n S d GYR VQE F T N LP LK MS LTQ YG C G R V E d i c TQ n A E C EH KP R L S AP F S VGK NY LQYAAV P T LLAGLGVS HA TY QCR QL LR P LQY H GDKWV LAP N a o AI L T I I R R EKP L d n i WS VDP i m AVL P I QP L GVP P KL TYTPKQLGRVA GQ F L S GI Q LLGKA E S S LKLS LA AI S VD LLP a TAI E QKKAEP EKDA AGAKWRAN P G F L I NKGG RKQI W L AVL P I Q GV c A QP TQP RS TQP DAYP M EYDS LKQQDTAE RPHR QP TQP R o F S n i DT D S AL L KLYL LP LP RK VQAAVYT RDS NH Y L RH LKVLYW I ARL FAS F S T DY RT Q T I GQGFKF LAVV RC WVMYI L MN P DAL LD QKMGPQE VA I L DNT HE K LEAR S S L K R N me a r W L T PI Q I I TP LWR HG P N P C F T LF DQQVI QF CC I T E F L D P P PAA KR R AETHDT VDQ EAGKN HLG S N W L PI Q I T G f L L A D Q I L R E o u s T I H P g o l ni c t si n oi s t 9 2 si 0 i d x 0 H 3 0 L t p i r c T D : g x a T D 1 . : n 2 e s R L T R L 3 s e D O 0 S - N O S 0 0 e l e r - 8 b p 2 a e Q : 9 8 6- T h t E S D I O N 1 1 1 2 1 6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000209_0001
u n S d i QE HF T AP S VN LP GK LQYLK MS LTQTYK C G QV AAV P T LLAGLGVS HA TY CE E T RQ HF P S V d c A S FNYL LR P LQYDKWVHLAP N L T I I K R EQ KP L S A FNY n a o d n P L i P i P P KQ F L T L YTPKG QLGR VA I QGKAI E S S LKLS E KKAGPAGAKWLA AI RAN WS VDP LKQL VL I QP P P F L I c m a A GQ S S G TQL L D LLPA QEKDA AYT P M EYDS LKP GLNKGGKQI L APGVGQ S GL QF QI DTAER RPDR QTQP RS TQD o L LP P LP RKQ ni NH Y L RHTV AAVYTKVLYWARL FAS P F S T DYL LP P LP RT Q I GQGFKF KMG LAL VV RC W I VMYL MN P DAL LDNH L RH PQE VA I L DNT EKI EAR S S L K R NYR T Q I me a r P NLWR P C F T LFQQVI QF DQ CC I T E F L D P P PAA KR R AETHH DTL N AGK S L Q VD L Q E HLGN WPI I T GP NL CW TR F f u s L A T Q I L R E T I H P P F L D o g o l ni c si n oi s t i 1 s H 3 i tsi d L t p i r : g x 0 H T D : g 1 . : n 2 e c s s e at- R L a O t- 3 0 0 el e D N r S N 0-8 b p 2 a e Q : 8 6 T h t E S D I O 3 N 2 1 -6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000210_0001
u n S d N LP LK MS LTQTYK C G QV EQE T LP d i c GK n A LLQ LRY P LAAV QYP T D LLAGLGVS HA TY KWVHLANT I I R R EQC P R LHF P S A F S VN GKQY NYLL LR P L a o n TYTP i QGKAKG I QLGR VA E S S P LL KLS A AKI VDP LKQ E L TYTP KKAGPAGAKWL RAN WS L I QP P P F LQGKA di m L LLPA QEKDAGLNKGGKQI L AVPGVGQ S GI L L LPA ca A AYT RKP M QEYDS KP QF I TAER RPHR QTQP S TQD L AYT P M AAV L YT QD KVLYWARLAS P F S D R YL P P LP RKQE o n TVG F L C I F N TL L H HTVGA i GQ FK AVV R W MYL M DA DN KMGQL E VA L DNTV EKI EARP S S KL RNY L R F RT Q I GQMG me Q a r QQVP I QF CC I T E F L D P P P I AA KR R ET DHH QDT EL AGKNS L HLGN W L PI Q I T GP NL CW TR FQK QQVP CT I F f A V L L A T Q I L R E T I H P P F L D C I E L o u s g o l ni c t si n oi 2 3 si si d t L t p i r x 0 H c T D : g 1 : n 2 e s s e R L a O t . - 3 0 0 el e D r S N 0-8 b p 2 a e Q : 5 8 6 T h t E S D I O N 2 1 -6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000211_0001
u n S d LK MS LTQTYGC GP P VW T N LP LK M d i c AAV n A QYP T D LLAGLGVKHA TYS EQE F P KWVHLANS TI I K S V R EQR RHA GKQYAAVT NYLL LR P L L QYP DK a o d n i KG I QLGR VA i KKAGP E S S P LL KL S A AKI C L S F KQL TYTPKGLG AGAKF L RAN WS P VDP P L P F LQGKAI QKA c m E a A QEKDAGLNKGGKQI L AVLQP QGI L L LPAE QK EKD YDS KP QF I TAER RPDR QP TI VP G S QD L AYT P MYDS K o AV L ni KFYT QD K LYWARLAS P F S G QRS TP LP RKQEAV L YT QLAL V EVVV A LR DC NW I TVMYF EKI L MN DT EARP S S AD LYL DLP NHRH LT Q TVGA F I GQ KFAL MGQL E VV me QF P P I AA THHTLGKNS L K L LNY RWRQK QVP I QF P A I a r D P KR R AE VD L Q LDEAQLVN WP RTP LT F QCTF D P RA f A T Q I L R E T I Q G N C L D C I E L P K R A o u s g o l ni c t si n oi 3 3 si si d t L t p i r x 0 H c T D : g 1 : n 2 e s s e R L a O t . - 3 0 0 el e D r S N 0-8 b p 2 a e Q : 7 8 6 T h t E S D I O N 2 1 -6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000212_0001
u n S d S LTQTYG C G P VW T N LP LAP LL HL d i c LAGLGVS HA TY QS A WVHLAP N REQE F P S VGK LT I I R R E Y KCRHAN LLQ LRYAYD P LQGLKWVAS GR V P ES n o R VA E S S LKLS LA AI S P L S F KQL TYTPKQKAGGA LG a d n i GP i AGAKWRAN WVVDP P L P F L I QGKAVK EKDA F I N T c m AG a A P LNKGGKQI L AP L I QP QGL L LPAE QDS LKPQDY QF QI DTAER RPHR QTGVP G S QD L AYT P MYVYTQ KVL C o K ni VV LR LY DC NWW I ARLAS P F S QRS LTP LP RKQEAFAL V RNW T TVMYF EKI L MN DT EARP S S ADY KLDLP NHRH LT Q TVGA F I GQ KLVV KMGQE F P A I L D A THH DQD me A ETHH DTL N AGK S LL L RNY RWRQQVP I QP R RAELLA a r VD L Q L AE T QH I L L DN WPI Q TP LT F QCTF DK AVLGD f R E T I I G N C L D C I E L P T W V A o u s L E P g o l ni c t si n oi 4 3 si si d t L t p i x 0 h T D : g 1 : nr 2 e c s s e R L a . O T- 3 0 0 el e D r S N 0-8 b p 2 a e Q : 9 8 6 T h t E S D I O N 2 1 -6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000213_0001
u n S d ANT I I K G QV E E T LP K MS LTQ YG C d i c P LLS R TY n A L CRQ HF P S VN GKQYL AAV AKWLA EQP L S A FNYLL LR P LQYP T D LLAGLGVS H KWVHLAP NT I I a o d n i KKG RA QN AG RK i PDI AKI VDP L L WS L I QP P P KQ F L T LQY GTP KAKG I QLG K R AE VA P E S S LL KLS AGAKWL R c m E a A W I AR RLAR AVPGVGQ S GI L L LPAE K EKDAGLNKGGK VMYF I L M S QTQP RN P P F S D RS TQD L AYT P MQ YDS KP QF I TAER YL P P LP RKQE RP AAV L YT QD KVLYWARL o n E i TK ELEA AGKNS DT ALDL NHRHTVGFKFAL V QLG S S KL RC W I R NY L RT Q I GQMGQL E VVL DNT VM EKYF I L EA me T TQI HLR N E LL Q TP LWRQK QVP I QF P A I A THHT LGK a r VT L AA L A E I L W T PI I GNCT F QCTF D P RAE DQDEAHL f I H P P F L D C I E L P K R A V L L A D Q I L o u s g o l ni c t si n oi 5 3 si si d t L t p i r x 0 H c T D : g 1 : n 2 e s s e R L a O t . - 3 0 0 el e D r S N 0-8 b p 2 a e Q : 1 8 6 T h t E S D I O N 3 1 -6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000214_0001
u n S d i A GTYQV C E RQE T LP K MS LTQTYK C HF AP S VN GKQYL AAV P T LLAGLGVS HA d c n A R R EQ KP L S FNYLL LR P LQYDKWVHLAP NT I I K R a o d n i A AI AN WS VDP LKQL TYTPKG QLGRVA E S S LL KLS i VL I QP P P F L I QGKA LA AI E KKAEPAGAKWRAN c m QI a A HL A R QP TGV QP GQ S GL L RS TQD LLP QEKDA AYT P M EYDS LKP GLNKGGKQI L QF QI TAER RPDR o AS P n i M R N F S T DYL LP P LP RKQ P D S S AL AAVYT D KVLYWARL FAS L KLD RNNH Y L RHTV RT Q I GQGFKF KMG LAL VV RC W I VMYL MN PQE VA L DNT EKI EARP S me NS L Q I TP LWR FQQVI QF P P I AA ETHH DTL N AGK S a r G RN E W T PI I HG P N P C F T L DQ CCTF D R DQ E HLGN f o u s I E L P K R A V L L A T Q I L R E g o l ni c si n oi 6 3 si tsi d t L t p i nr x 0 H c T D : s R L g a 1 t . 3 : 2 e s e D O - 0 S N 0 0 e l e r - 8 b p 2 8 a e Q : 3 6 T h t E S D I O N 3 1 -6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000215_0001
u n d GYQEQF TVNK P L TS LTQTYV GYQ d i c TQ n A E CRHAP S YGLQ RYAA YV P LLAG HLGVEG T TQC a o AKI P L S P FNQL L P L PQGDKWVA L S AP N LLKK E AKP d n i WS V i VLD I QP L P P KL TYT QF AKI QLGR GV ES LKE EGK I S VL GL I QLG c m A L K PAE KKAA P GAGAKGNG S WVI a A QPGV P T P G S Q DRS S LT PQL P D LYT PAKP MQEK EYDS D LKP F L I NKGKDK A TQQDTA QPG Y EYDE P T S Q o F T ni DS ALY LKLDL NH L L RHR TQ TVQ GAAV FYLKVL C W I AKY F T L F T D L a LR N T Y P R LWI GQ FK A RQKMGPQL VV LRNW TVM QE F V P A I AD THHEK LKFM DS AL VS I Q L K R me r WP Q I GNCT F QQV TI F D P RAE DQDT EATML L W L P Q I f T I I H P P F L D C C I E L P K R A V L L A T Q V K E T I I H o u s g o l ni c t si n o 7 i o t 6 8 si 7 o t si d t p i x S - 5 H c : g S: x L t : nr 2 e c T s R L n s e u a g a T R 1 . 3 D B 2 r T T - T 0 N - 0 0 e l e r 4 C -8 b p 2 a e Q : 5 7 8 6- T h t E S D I O N 3 1 3 1 6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000216_0001
u n S d V EQE F TVN LP LK MTS LVKE GYQV EQE F T N LP L d i c n A R LHS AP S K YGLQ RYAAV P LLAGGK TQCRHAP S V K YGLQYA a o P FNQL L P L Y PQGDKW RVGV RP E AKI P L S P FNQL LR P L PQ d n D i QP L P K i VP Q F L T QYTAKI QLGGVG GL I LG c m L K PAE KKAA P GG TWA S V VD W LDP L P KL TYTAKI K VI QP Q F GL I QLG L K PAE a A P G RS S LT PQL P LYTMQEKS D KP FGKE A QP TGVP G S S QLLYTMQ LD PAKP QEYD VL TQQDK I S P S Q DR LT P P D PAKP QEY o YL H n i D RHR TVGA NN FA KFYLKVL CKV A F DTLYL H L RHR TVGA FA K YLT Q I GQMGQLA EVVV LR DNS I G S ALD NNLT Q I GQMGQ me a r T GP R LWRQK QVP I QF P A I A THDR L K LR QT Y P R LWRQK QVP I Q P N P C F T LF DQ CC I T E F D P RAE DQVG WPI I GNCT F QCTF D f u s L P K R A V L L E T T I H P P F L D C I E L P o g o l ni c n 7 o 7 si 7 7 o s 7 t si oi si d t t pi S- 9 4 H o t t c : S 5 1 i Ho t g S: x T - 5 c : L t g S: 1 : nr 2 e c s L s e D B n u a g a R L n u at g a . 3 2 r T T - T N - B r T - t- 0 0 e C 2 N C 0- l e r 4 4 8 b p 2 a e Q : 8 T h t E S D I O 9 N 3 6 1 -6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000217_0001
u n S d K M i AV S LTEN G P T LL QVQE AGGDG S TY CEHF T P S VN LP K MS LTQTY GKQYLAVT LLAGLGV d c n A YDKWVHKDK EQP R L S A FNYLL LR P LA QYP KWVHLAN a o G d n L i Q GR KAGVA P EYYE AKI VDP LKQL TYTPKGD LGR VAS S P LL i AKT L WS L I QP P P F LQGKAI QKAGP E K AGAK c m K a A EKDAG DS KP L F F S M AVPGVGQ S GI L L LPAE K EKDAGLNKG QF I KI Q QTQP S TQD L AYT P MQ YDS KP QF I TAE o V L n FYT QD K LV TML P F S D R L P P LP RKQEAV L T QD K LYWA i LAL E VV VVLR DC NVKL S GE K DT S ALYL KLDS NHRHTVGA FKFY AL VV Y L RT Q I GQMGQL RC E W I M VVL DNTV EK me F P A I A THGV RP LLR QTP LWRQK QVP I QF P A I A THHTL a r P KR RA AE VD L QG A WPI I GNCT F QCTF D P RAE DQDEA f u s L G W D T I H P P F L D C I E L P K R A V L L A T Q o g o l ni c t si n oi 7 si si d t 3 pi x 0 H T D : 1 L t : nr 2 e c s s e R L g a . O t- 3 0 0 el e D r S N 0-8 b p 2 a e Q : 8 6 T h t E S D I O 1 N 4 1 -6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000218_0001
u n S d i K S C HA GTYQV C E RQE HF T L AP P K MS S VN LTQTYK C GK LQYL AAV P T LLAGLGVS HA d c n A T I I K R EQ Y KP L S FN L LR P LQYDKWVHLAP NT I I K R a o L d n S i WLA AI VDP L i RAN WS KQL TYTPKG QLGR VA E S S LL KLS VL I QP P P F L I QGKAI KKAGPAGAKWLA RAN c m G a A RKQI RPDL A R QP TGV QP GQ S GL L LPAE QEKDAGLNK RS TQD L GGKQI L AYT P MYDS KP QF I TAER RPDR o RLAS P n i YF I L MN F S D DT YL P P LP RKQE EARP S S ALDL KLNNH Y L RHTVGA FAV L KFYT QD K AL VLYWARLAS RT Q I GQMGQL E VVV LR DC NW I TVMYF EKI L MN ARP S me GKNS LLR QI TP LWRQK QVP I QF P A I A THHTLE GKNS a r H I L L G RN E W T PI GNCT F QCTF D P RAE DQDEAHLGN f u I H P P F L D C I E L P K R A V L L A T Q I L R E o s g o l ni c t si n oi si si d t p i x H T B : L t : nr g 1 2 e c s s e R 2 4 at . - 3 0 0 el e D N r 0-8 b p 2 a e Q : 3 8 6 T h t E S D I O N 4 1 -6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000219_0001
l c q M u e S GLPNY FDGG TTGAGAP G TKGI GP S VWW EL TDLE R L LLKAMKS NR TE S TYGP CA n d i TYQ CE RQ HF P S VNKQ P YLAVT LLL Q K AGLGVS HA d c n A EQP L a o AKI S A FNYGLL LR P LA QYP S VDP DKWVHLANT I I K R P LKQL TYTPKG QLGR VA E S S P LL KLS LA d n i W ic m AVL P I Q GVP P QF L I Q a A QP T P G S G LGKA PAI E KKAGPAGAKWRANI S QRS LT PQL P D LL YT PAKP MQEK EYDS DA LKP G F L I NKGG RKQL TQQDTAE RP LDR o F T D n LYL H L i DS ALDNL RHR TQ TVQ GA FAV FYLKVLY C W I A MR FAS N I GQ K MGQLAVV LR W DNTVKYI L M ARP me L K a r W LR N T Y P R LWRQKVP I QE F V P A I A THHE TLEKNS S T PI Q I I HGNCT F QQ CTF D P RAE DQDEAG HLGN f u s P P F L D C I E L P K R A V L L A T Q I L R E o g o l ni c t si n o si si di t L t p i r x L H c T B : g 1 : n 2 e s s e R 2 4 at . - 3 0 0 el e D N r 0-8 b p 2 a e Q : 5 8 6 T h t E S D I O N 4 1 -6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000220_0001
u n d GYQEQF T N P L S LTQTYK CA G d i c TQ n A E CRH KP L S AP F S VGK NY LQYAAV P T LLAGLGVS Y QL LR P LQY V K TY Q GDKWV LAP N a o AI L T I L I R EK d n i WS VDP i VL I QP L P P KL TYTPKQLGR VA Q F L I QGKA E S S LK S LA AI S AI E KKADPAGAKWRANI WV c m A a A QPGVG S GL L P TQP RS TQP D LLP AYT P MQEKDP EYDS LKP G QF L QI NKGG DTAERKQ RPDL A R QP P T o F S n i DT DYL S AL L KLDLP LP RK RNNH Y L RH VQAAVYT LKVLYWARL FAS F S T RT Q T I GQGFKF KMG LAVV RC W I VMYI L MN P DA PQE VA I L DNT HEK LEAR S S L K me a r W L P Q I TP LWR FQ QQVI QF P PAA ETHDTAGKN GS L f T I I HG P N P C F T L D C C I T E F L D P KR R A VD L Q L AE T QH I L L R N E W T PI I o u s g o l ni c t si n oi s s t i i d x G H x L t p i r c T + : : n 2 e s R A0 g a T 1 t . s e - R 3 0 D 5 N 0 0 e l e r - 8 b p 2 a e 8 Q : 7 9 6- T h t E S D I O N 4 1 4 1 6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000221_0001
u n S d P VW T N LP LAP LL HLANT I I KGRA G Q d i c S n A R CE RQE HF P A S VGK NY LQYAYDKWVAS P LLLS RG Y WD TQC QL LR P LQGLGR V P E AS GAKWLA a o RAGVK EKP d n P L S i VDP F i P L P K F L T L YTPKQ I QGKA AV E KKAG EK GL S DA P F I NKKG TAG ERKQNI KE RPDL K S AI V WS V VL I c m L a A I Q GVP QG P G L L LLP S S QP DAYT QDLKQQDLYWARLAR I KP M EYVYT LKV RC W I MYF L M S KA APG o QR LT P LP NS G QP TQ ni D LYLHRHRVQAAF LD NNL T Q T G LAVVDNT V E KI EARP I DR F S T D L I GQ FKE V P A QKMGPQF I LTHHT LGKNS S VG T D P RAA E DQ LDEAQ I L L G RNEKK S A LKL R me a r RT Y P R LW TR F QQV TI FQ KRAVL L AT TQH ELGK L P Q I f u s Q I G N C L D C C I E L D P T W V A L G E D P VT L AA L A E I L E E N G W T I I H o g o l ni c n - ) L s 7 t si o si di t p 4 3 3 i 1 H o t S x L t i r c DK ( : g : a g a T : n 2 e s s e L O7 R 1 . o t t - 3 0 N T - 0 el e D r S S C 0-8 b p 2 a e Q : 8 6 T h t E S D I O 1 N 5 1 -6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000222_0001
u n S d V EQE F T N LP LK MS LTQ YG C GV P G QV E d i c n A R LHS AP F S VGK NY LQYAAV P T LLAGLGVS HA QL LR P LQY RGRA TY QCR GDKWVH a o A LAP N L T I L I RGWD EKP L d n DP i QP L i P P KL TYTP QF AKI QLGR V E S S LK S WLAGVK AI S VD GL I QLGK PAE KKAG APAGAKG RANI KE S WVL I Q c m V a A P G S RS LT PQL P D LL YT PAKP MQEK EYDS D LKP G F L I NKG RKQL K I V APGVP TQQDTA Y E RP LHRKA QP T S QR o YL H L i RHR TVQ GAAV FYLKVL N C W I AR FAS NS I G R F T D LY n DNLT Q F I GQ K A MGQL VV LRNW T VM KYI L M RPDG DS ALD me a r T Y G P R N L CW TR FQK P QQV QE F V P A I AD THHE TLEANS S V T L K R N T f P P F L D C C I T I E F L D P P KR RA AE VD L Q LD AEAGK D QH I L L G RN E E L EKK G K W L PI Q I G o u s T I H P g o l ni c t si n - ) oi 5 L 3 si 7 o t 7 o si d t p 2 i 0 1 h: S: x t S L 3 L t r c DK ( g at g T : n 2 e s s e L a R - 1 1 . 3 O7 o t - N t- B 2 K 0 0 el e D r S S C 4 0-8 b p 2 a e Q : 8 6 T h t E S D I O 3 N 5 1 -6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000223_0001
u n S d QE F T N LP LK MS LTQTYK C GV P G QV EQE d i PV KQ AVT L G GVS A R TY F c H n A S A F S G NY L YA P L A L H G A QCRHA QL LR P LQY GDKWVHLAP N a o L T I L I K RGWD EKP L S F d n P i P L i P P KL TYTP QF L I QGKAKQLGR VA E S S LK S P LAGVK AI S VDP AI E KKAGPAGAKWRANI KE S WVL I QP L P P Q c m a A G S S GL L LL T o LT LPQP DAY KP MQEKDA EYDS LKP G QF L QI NKGG DTAERKQ RPDL K R I V APGV KA QP T P G S S QRS LT H LP ni N YL RHRVQAAV FYT LKVLYW I ARL FAS S I G F T D P LYL T H Q T G GQ FKLAVV LRC W TVMYI L M RN PDR D a P R I QKMGPQ S ALDNL N QE V P A I AD TN HHEK LEANS S VG T L K R NYR me r LW P C F TR LF DQQV CC I T I E F F LD P P KR RA AE VD L Q LDT AE TAGK QH I L L G RN E E L EKK L G K W T PI Q I I T GP NL C f o u s H P P F g o l ni c si n oi si 7 o 7 t o t s s S i t i d t p H S x - L H L t i r : c g : a g T L 3 1 : g a 1 . : n 2 e s s e t- a R K t 3 0 N T- B - 0 el e D r C 2 4 N 0-8 b p 2 a e Q : 8 6 T h t E S D I O 5 N 5 1 -6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000224_0001
u n S d i T N LP LK MS LTQTY c P S V K CAGV P G QV GK LQYAAV P T LLAGLGVS HKGRA TY QC E RQE HF T AP S V dn A NY QL LR P LQY GDKWVHLAP N a o L T I L I RGWD EKP L S FNY Q d n i K F L T L YTP i I Q AK LGK PAI QLGR VA E S S LK S E KKAGPAGAKWLAGVK AI RANI KE S WS VDP VL I QP L P P KL Q F L I c m G a A QL P D LL AYT KP MQEK EYDS DA LKP G QF L QI NKGG DTAERKQ RP L K I V APGV LDRKA QP T P G S GL o LP S QRS LT PQP D n i RHRVQAAV FYT LKVLYW I AR FAS T Q T G NS I G F T D LYL H LP H I GQ FKLAVV QKMGPQ LRC W TVMYI L M RPDR DS ALDNL R T Q QE V P A I AD TN HHEK LEANS S VG T L K R N a T Y P R I me r W TR LF DQQV CC I T I E F F LD P P KR RA AE VD L Q LDT AE TAGK QH I L L G RN E E L EKK L G K W T PI Q I I LWR HG P N P CT F f u F L D o s g o l ni c t si n o 7 o si di t - L si 7 p t S x V3 1 H : o t S L t i r c : g T : n 2 e s s e a t- R B K 2 4 7 g o a : g 1 . t T- a t - 3 0 0 el e D r C S N C 0-8 b p 2 a e Q : 8 6 T h t E S D I O 7 N 5 - 1 6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000225_0001
u n S d N LP LK MS LTQTYK C GV P G QV EQE F T N d i c GK n A LLQ LRY P LAAV QYP T LLAGLGVS HAGRA TY QCRHAP S VGK L GDKWVHLAP N a o L T I I K RGWD EKP L S FNY QL L d n TYTP i Q i GKAKQLGR VA E S S LKLS P LAGVK AI S VDP AI E KKAGPAGAKWRANI KE S WVL I QP L P P KL TY QF L I QG c m L a A LL AYT QEKDA P G F L I NKGG RKQL K I V APGVG S GL L LL KP M EYDS LKQQDTAE RPDRKA QP T o RVQAAV S QP RS TQP DAY K n T G FYT LKVLYW I ARL FAS S I G F T D LYL LP HLP RV i GQ FK MGQLA EVVV LR DC NW TVMY EKI L MN ARP S DR D VG S ALD NNL RH TQ T I GQ QK P P A I THH LE N T L K R YR QK me a r QQV TI FQF P RAA E DQDTAGKG S EKK L P Q I TP LWR F QQ f u C C I E L D P K R A V L L AE T QH I L L R N E L E G K W T I I HG P N P C F T L D C C I o s g o l ni c n s 7 7 t si oi s t i i d x H o t o S t S 9 2 5 L t p i nr c T : s R g a : t g - at Lc n u 1 . 3 : 2 e s e - - B r T 0 0 el e D N r C 2 4 0-8 b p 2 a e Q : 8 6 T h t E S D I O 9 N 5 1 -6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000226_0001
u n S d L i Q P YLK MS LTGVK G QV AAV P T LLAGLKE TY CE RQE HF T AP S VN LP GK LQYLK MS LT AAV P T LLAG d c A R P LQY GDKWVH AL K S EQ KP L S FNY QL LR P LQY GDKWVH n a o d n TP i KAKQLGR V E S S I V A AI S VDP i PAI E KKAGPAGKS G WVL I QP L P P KL TYTPKQLGR VA Q F L I QGKA E AI E KKAGPA c m a A T P MQEKDA EYDS LKP G QF L QI N I R APGVG S GL L DTD Q LLP T QEKDA P G F L I VG P TQP RS TQP DAYP M EYDS LKQQD o QAAV n GFKFYT LALKVLY VV RC WE T F S T D LYL LP TLK H LP RK VQAAVYT LKVL GK DS ALDNL RH TQ TQGFKF LAVV RC i MGPQE V P A I L D TNHE ENK L K R NYR I G QKMGPQE VA I L DN me r VI QF P RAA E HDGDG S L P Q I TP LWR F QQVI QF P PAA ETH a T E F L D P K R A VD L Q L A K D K W T I I HG P N P C F T L D CTF D R DQ f u s C I E L P K R A V L L o g o l n 7 7 i c t si n oi s t o t S 4 3 si 5 H o t i d S L t p i nr x c T - s R Lc : n g : u at g 1 a . 3 : 2 e s e D B 2 r - t- 0 0 e 4 T N C 0- l e r 8 b p 2 a e Q : 8 6 T h t E S D I O 1 N 6 1 -6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000227_0001
l c q u e G I A MLPNYWFDGGR TTGAGAP RGS I A n S EK G di QS L I G GP S VWEL TDLEL L P LKAMKS N DR TYQ LTE K QT S I G QC E RQ HF AP S VN GK LQYL AAV P T LLAGLGDR d c n A L G a o S VT EKP L S FNY QL LR P L Y n S E PQGDKWVH LK A A L S APVG T K I S VLDP P L P KL TYTAKI QLGR GV ES L E LKK d i i G c m NEG ENK W AVP I Q GVP GQ F S GL I Q LLG LL K P a TAE QKKA P EKDA P GALGA NKEG ENK A T DG Q YG TQP S Q Y M S K F I TA DG KDS K P F S D R LT P P L D PA RKP QEYD VL TQQDYWGDS o n i W TYYE KT DT F L S ALY DL NHRHTVGA FA KFY ALK VV RL C W I K YYKE LKL RNY L RT Q I GQ KMGQL E VV AL DNTV EKT F L me H DF KS I M L Q I TP LWR FQQVP I QF P P I AA ETHH DTF S I M a r A V MQ L W T PI I HG P N P C F T L DQ CC I T E F L D P KR R DQ EK Q f o u s A V L L A T V M L g o l ni c t si n o 7 i 7 s t o t 0 i d x S 6 s 5 i h o t S L t p i nr c T - s R Lc : n g u a : t- g 1 a . 3 : 2 e s e B r N t- 0 0 el e D 2 r 4 T C 0-8 b p 2 a e Q : 8 6 T h t E S D I O 3 N 6 - 1 6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000228_0001
u n d GYQEQF T PVNK P L TS LTQTYKV GYQ d i c TQ n A E CRHA S YGLQ RYAA YV P LLAG HLGV A S EG T TQC a o AKI P L S P FNQL L P L PQGDKWVA L S P N L T LLKK E AKP d n i WS V i VLD I QP L P P KL TYT QF AKI QLGR GV ES LKWE EGK I S VL GL I QLG c m A L K PAE KKAA P GAGAKGGNG S WVI a A QPGV P T P G S Q DRS S LT PQL P D LYT PAKP MQEK EYDS D LKP F L I NKG RKDK A TQQDTA QPG Y E RYDE P T S Q o F T ni DS ALY LKLDL NH L L RHR TQ TVQ GAAV FYLKVL C W I AR KY F T L F T D L a LR N T Y P R L I GQ FK G RQKMPQLAVV LRNW TVM QE F V P A I AD THHEKY LI EKFM DS AL VS I Q L K R me r WP Q I GN W TF QQV TI F P RAE DQDT EAGTML L W L P Q I f u T I I H P P C F L D C C I E L D P K R A V L L A T QH I V K E T I I H o s g o l ni c t si n o 7 o si di t t 3 1 si 7 o p i x S - 6 c h: t S: x L t r c T : n 2 e s R L n g at g T 1 a R . 3 s e D Bu 2 r - T N t- 0 C 0 0 e l e r 4 -8 b p 2 a e Q : 6 8 8 6 T h t E S D I O N 6 1 6 1 -6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000229_0001
u n S d V EQE F T N LP LK MTS LTQTYKK I V GYQV EQE F T d i c n A R LHS AP S V K YGLQ RYAA YV P LLAGLGVS KA TQCRHAP S V Y a o P FNQL L P L PQGDKWVH AL S AP N L T L S I G E AKP L S P FNQ d n D i QP L P K i VP Q F L T QYTAKI QLGR GV ES LKWDR I GL I LG c m L K PAE KKA P G S VLDP L P KL K AGAGAKGV ET W AVI Q VP Q F GL I a A P G RS S o YLT LPQL P LYTMQE S D KP F L I N TK AG RLKK QP TGP G S S QL HL D PAKP QEYD VL TQQDYW E AREGK P S Q DR LT P P D P ni D RHR TVGA NN FA KFYLKV Y LT Q I GQMGQLA R L C E W I MR YE GNG S F DTLYL H L RH VVV L DNTVKI KD A DK S LD NNLT Q I me a r T GP R LWRQK QVP I QF P A I A THHE TLE GYYE L K LR QT Y P R LWR P N P C F T LF DQ CC I TF D P RAE DQDEAHKT L WPI I GNCT F f u E L P K R A V L L A T Q I F F M T I H P P F L D o s g o l ni c n 7 o 9 s 7 - L s 7 t si oi si d t t pi S- 0 i 6 c H o : t 3 1 i h o t L t g S: x S T G + K : g : 1 : nr 2 e c s L s e D B n u 2 r at g a R A7 at g a . 3 T - t N - 0 C 5 o t - S N T- 0 0 e C 0- l e r 4 8 b p 2 a e Q : 8 6 T h t E S D I O 0 N 7 1 -6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000230_0001
u n S d LP K MS LTQTYK C V P L QNK RF d i N c GKQYLAVT LLAGLGVS YAG GRA K S GS WNYWLD D n A LL LR P LAYP KWVVLANT I I KGWD TMMP VWE T PV a o d n TYTPQ i Q i GKAKGD I QLGR VAS KKADP ES P LL KLS R AGVK EGP S EQF S Y AGAKWL RANKE HGYQRHS ANQ c m L LPAE a A L AYT EKDP G P MQ LNKGGKQI L K S L RT EQC P LP F LKL YDS KP F I TAER PDR I V K HAKI VDP P F L I o RK n TVQE GA FAV L TQQD K i GQ KFY LYWAR RLAS S A G E S L AL MGQL E VV VVLR DC NW I TVM KYF I L MN I R DWVI Q VL QGL GP G ARP S DG E I AP TQRS S QP D LT P LP e QK QVP I QF P A I A THHE TLE GKNS V ET NQP S DYLHRH m a r Q CC I TF D P RAE DQDEAHLGNLKK L F T LDNL TQI f u s E L P K R A V L L A T Q I L R E E G K T D A L N Y R W R o g o l n n T i c t si oi si d t p i x W T - V L t r c : n 2 e s s e R L 1 . M 3 0 0 el e D r M 0-8 b p 2 8 a e Q : 6 T h t E S D I O N 2 7 - 1 6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000231_0001
c q e YQF K P W VGQRT DQQS EHEGI P WNY LDG L GR T L TGA u n S d d i c G L G E E R T A LL T LGAGAA P P KA GADKGT I GYP S VWW E TD VNE L P L LKAMT MNNR TE S TT GPATQR EQF P S K YGL Q RYAA YV P L n A NK a o LQ P YLAVTS n GLR P LAYP LLL AGQGY HL VE LANS C T H I AE I KAKI C R S P LHS A F NQL L P L P QGDK VDP L K F L TYTAKI QLG d i L ic m TYTPQ KGD LKW RVAS P LL L S RW AVLQP P GL I QL G L K P AE KKA a A QG L L KAI QKG V ADP ES AKL L RA AQP T I VP GQQL LYT MQEKS D o LYPAE QK EK A DP GLGKKG RK G QP S QP S S T P D P AKP EYDLKT n KT P M EYDS KP QF I D TAG E RPDF T i A D R YL P L RHR TVQ GAAV F YL me R TVQ QGA FAVL KFYT QD K a r GKMGQLAL VVV RLY C W W I ARL FADS ALDL NHT Q VMYI L M L K L L NYL RWI GQ F K G RQKM QLAV VP I QE F V P A I Q Q VP EVA L DNT EK EAR WPI R QT P L T F QQ C T F D P RA f u s I Q F P I A T H H T L G K N T I I G N C L D C I E L P K R A o g o l ni c t si n oi s 3 3 si d t p i i H 3 : DG 3 G D DD 1 L t : nr 2 e c s s e x g a L t O V L O V . 3 0 D T - S S 0 0 e - l e r R N 8 b p 2 a e Q : 8 6 T h t E S D I O N 3 7 - 1 6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000232_0001
u n d S L QG G CA GYQEQF P VNKQ P LAVT S L Q d i c LAG n A W HLAV a o RVA L S P N S L T H I L I S R T R EQC RHA S YGL RYAYP L LAG HLG A AKI P L S P F NQL L P L P QGDKWVA L S P d n i GV i A P E S LKWLA S VLDP L P K GAGA KKG RANI W AVI QP Q F L TYTAKI QLGR GV E S L GL I QL G c m P F L I N L K P AE KKAA P GAGA a A QQDTAG RKQL QP GVP G Y E R P LHR S P T S QR S S L T P QL P D LYT P AKP MQEK EYDS D LKP F L I NK TQQDTA o KVL n i V L R C W I AR F A N F T D LYL H LHR TVQ GAAV F YLKVLY AD TNW T VMYI L M R P DS ALDNL R T QGQ F K G LAVV HHE T K L EANS S L K R NYR I QKMP QE V P A I L R C WW I AD TV TN HHE me a r E VD L Q LD AEAGK D QH I L L G R N E W L T PI Q I I T P LW HG P N P C F T R L F DQQV C C I T I E F QF L D P P KR RA AE DT VD L Q E f u L A T o s g o l ni c t si n o si B - 2 1 t si di t : p x H: g a 0 9 4 a D L t i nr c T s R g at T - t a T D A L _ 0 L 9 1 . 3 : 2 e s e - C D- B 0 0 el e D N r C 2 4 0-8 b p 2 a e Q : 8 6 T h t E S D I O N 4 7 - 1 6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000233_0001
u n S d Y VK CAGE RDHS d i c N S T H I I KGL TQ QS n A a o L KL S R S WLAGKS P S d n i i KG RANI GG c m KDH T a A G E RK AR P Q LDL AR S G P T E L E o R F n M NKKR L i KYI L M AR P S RG GGA me a r L E AGKN QH I L L G S NS GTD DV f u s R E P Q A A o g o l ni c t si n o si di t p L t i nr c 1 s . 3 : 2 e s e 0 0 el e D r 0-8 b p 2 a e Q : 8 6 T h t E S D I O - N 6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000234_0001
u n S d i R AGQL L LM GG LA LKL TDS R GKC I HP DGE F T VR L D E D L E AA S AWG T LQP VT L TAAAKG N D d c n A S DVG AAAP VV P E L R R L G P AT P L a o P ML T L K R RY K KQP T d n i G P QAH LKT NR T D R N i KR E P GQ F L LYQF K P S KY A L VWQR AD P QQR I HE WLDGG R R T TGGGA KP R GAS DE KGT I c m R RMP P MLWP NWE TDL E KL L P LKAMS N L T E S T TGP A a A GR RD RAGP S V EQF P S VN GL QYL A AAVT LAGQGYK CA o Q ni AG L S S S AG I S S TYQ EQ KC P R LHS A P F NYL L R L KQL TYP L T P QYP L WVH AKGDK AL LAVS H I I K GR V E S P L NT L S R me a r T G KP LNAI VDP P F LQG AK R R S EHWS L A AV P I QP QGI GV P G S S T Q P L L DL L KAI YP EQLAGP AS AL KWLA AN QK A KT P M E YEKDA DK S KP G QF L QI GKK DN T A G R KQI L f u s G R P D R o g o l ni c t si n o - s B 2 si di t i p x H0 9 4 - L 1 B 2 L t i : nr 2 e c T : s g t a T A L t a 4 1 . s e R at- D D- D_ 0 3 9 0 0 el e D N r N 0-8 b p 2 a e Q : 8 6 T h t E S D I O N 5 7 - 1 6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000235_0001
u n S d K d i RGAGQGV P H I KGH c GG TDS GR VL L n A S R L D E DME L A a o d n i GDV P QAAA L i KR E H LA KP T V L R V P DE c m RMP GQNKT N YR F a A R GRDP P L GS LW o QR P NWWL E T GR S A AG n T P YS VQF P QEHA S i A L S S G I S L S EQC R L S P F N K me T a r KP RNAK HWI P AK R S E A AS VD L LQP P P QF G f u s V I V G S Q o g o l ni c t si n o - i si 0 9 s - 1 0 0 9- 1 0 i d t L t p i : nr x c T H: 0 9 t T D T D 1 s R g at a AL AL . - D D- O D- O 3 0 2 e s e D N S N S 0 0 e l e N r - 8 b p 2 8 a e Q T h t E S D : I O N 6 6 7 - 1 6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000236_0001
u n d AS LKT GKC I H D E F KG GP S VWE P VNK P d i c A n A L AWGG LQP VT L L TA L R R L P A a o P T L P L P M L R T KAANKD D R RYKQP TY T EQQEQF S YGL Q RY A C RHANQL L P L P d n i LYQF i DGG R K E R T A TVWQ GGGAA P QQR KP R GAS I D EHE KI P L KGT I WS VLDS P F L KL TYTA QP P F L I QL GK P A c m DL a A VNKL L P LKA AMS N L T E S T TGP A AVP I VP QGL L L T GL QYL M AAVT LAGQGYK CA QP TGP G S S Q T P DAY KP E o Y n QL L R i L TYP L T P QYP L WVH AKGDK AL LAVS H GR V E S P L N L T I L I K F S QR S R L P L P HRVQA A DT AD LYLHR T Q TQGF me LQ a r I L L G DL L KAI YP EQLAGP AS QK A KT P M E YEKDA DK S KP G QF L QI GAKWLAN S L KK KLDNL I G DN T A GG R QKMG R K P QI P DL L R N R W T P I Q I T Y GP R LW T R F QQV C T I F f o u s N C L D C I E L g o l ni c n s 0 t si o si di t i x H : g 0 9- 1 T 0 L t p i r T : g a T9 t a AD L 1 . : n 2 e c s s e R at- - C D D- O 3 0 0 el e D N r C S 0-8 b p 2 a e Q : 8 6 T h t E S D I O N 7 7 - 1 6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000237_0001
u n d LAVT LGQGVK CAGRDH R GAS R L DDML S d i c AYP L LAHLA n A QGDKW RVA L S P N S L T H I L I S K R G S L TQS Q S G TDGVR AL E L E P AA L AW S P S DVAAKP T VVD E NL R R a o d n KI QLGGV P E S LKWLAGKGS G P QA E H L QNR KT R F LYQ i i E KKA c m K AGAGAKG RANI G KDH T KR RMP P GS WNYWLDG L G E a A QE YDS D LKP F L I N TQQDTK AG RKQL Y E GT L R RDP MMP VWE T P DNK AR P LD AR S P E E GR RAG GP YS QEQF S V YGL o AV F YLKVL C W I MR F i E MNKK GR L QG S S ATQC RHS A F NQL L n K QLA VVV A L R DNW TV E KYI L EAR P S R GGA A L S GI S L S E AKI P L VDP P L P K F L TY G me LQ a r Q DF P P P KR I RAA AE T VDHH L Q LDT AE L N TAGK QH I L L G S R N S E G TD P DV T KP N KR S HWS VL I Q VP GQS G QI L L L L Y f u s Q A A A R E A A P G P S T P D A K o g o l ni c t si n o si di t 0 9 1 -si B 2 0 9 pi -T 0 x H: 0 9 4 T -TB L t r c AD L T : n 2 e s s e D- O R g t a a A A2 4 1 . 3 S t- D D- D - 0 0 el e D r C N N N 0-8 b p 2 a e Q : 8 6 T h t E S D I O N 8 7 - 1 6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000238_0001
u n d KT G C d i c G T LDAE F KG R GAS R L L DDML S KT n A L G LQP a o F P A P T L P VL P ML T L R T KAANKD D R RYK P G TDGVR L E L E P AA I QT S DVAAAP T VVD E NL AWGG L R R L F P d n i R K i R T AVW L TGGQAA P QQR S EHE G P QA E H LKNR T R F LYQR K AGKP NR GAS DKGT I KR P GQ S WK NYWLDG c m L P L L G E R T L a A Q LKAMS L T E T TGP A R R RM DP P ML P VWE T RYAAVT L LAGQ P DNKL P HLGYK C HA GR RAG GP S EQF S VGL QY on P L i T P QYP KWV AKGD A LAVS GR V E S P L N L T I L I K Q S R G S ATYQ A AS S I S S EQ KC P R LHS A P F NYL L R L KQL TYP L T P A me a r K P AI T P M EQLAGP AS EQK YEKDA DK S KP G QF L QI GAKWLAN L KK DN T A GG R R K P QI DL T G R KP LNAI VDP P F LQ AK R R S EHWS L A AV P I QP QGI GV P G S S T Q L L G L L KA YP T M f u s P D A K P E o g o l ni c t si n o - s di t i 0 p x H0 9 T AG 9-TG si L t i : nr 2 e c T : s g t a D + + A AA 1 . s e R at- D- N0 5 D- 0 5 3 0 e l e D N N 0 r 0-8 b p 2 a e Q : 8 6 T h t E S D I O N 9 7 - 1 6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000239_0001
u n d i G C H c LQP VT LD L TAE F KG R GA n KAANKD P G TDS GR VL R L L D ML E D S K L E T G P AAAWGG LQ d A A a o P T L P L P M T L R RYKQT S DVAAAP VVD E L R R L P A P T L d n i A i TVWQR AD GGGAP P RQQR I GAS DEHE GQAH LKT NR T R N F L LYQF KAV KGT I P KR E P GQ KYWLDGG R E R T TG c m a A L LKA AAAMK T S N T E S L L T TGP A R RMP VS I W P NWE TDL L L P LK AGQGYK CA GRDP AGP S V EQF P S VNKL QYLA o QYV n GP L KWVV AL LA i I QLG P VS Y T I I K R QR GR S S AGTY QR C RHS A F NYG QL L R P LA P QY K D R DV P E S N S AS I S AS L L L LA L GL S E AKI P L VDP P L P K F L TY QG T G KAKI Q me E QK EKA DP GLGAKWRAN I T P RN S LQP QGL I L L P AE K a r Y DK S KP QF QI KK DN GKQL KKS HWV I VGS Q L LYT MQE f u s T A G R P D R A R E A A P G P S T P D A K P E Y D o g o l ni c t si n oi 0 3 si d t 9 - 3 L t p i T DG D 1 : nr 2 e c s AL s e D- O V . 3 0 0 el e D N S r 0-8 b p 2 8 a e Q : 6 T h t E S D I O N 0 8 - 1 6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000240_0001
c u e n S d L G P QGL ADH d i C I T K S AI VHKG A LAL LG P c P P VH n A L T LDAE F KGD RGQGP I LMH ML T L KAANKP GG R L RLD E DE LAS KT G AA L AWGG LK AQCI H TP TL PVLML WP R T D R RY RKI QT S S VAL P V LP E L RR QL F P P LWP RT a o d n i GQAA P P QQS EHE G PGAKTR V T DN LY GR K T AVGQAA P i AGKNR c m A AMS T GA E S NKGT I KR RL GQNKYRFDG LE RL TGAGKN L LGQT TGP CA R RMS WNWWL TDNKLP L LKAMS L a VT L AHLGYG S HA GRGL P P S VQE F PV YGLQ RYAAVT LLA o P KWVA LAP VT I I R R QGGYQE RHA S QL L YP L PQY GPKWV ni D LGR GV KAA P E S GAS L N L L S A A GAKWL RA L S T G P T EQ KC I P L S FNL T DP L P K F L I QLG T L KAK AI EQDGR V K L G KA P me a r K S D KP QF L QI DN TKK A GG R K P QN HI L K AKA R WS VQP VL P QGLLYP TMQE KDA P G F f u s I V G S Q D A K P E Y D S K Q Q o g o l ni c t si n o 6 si di t p 3 t L L t i : nr a c B se D- 2 4 1 . 3 0 2 e s el e D N 0 r 0-8 b p 2 a e Q : 8 6 T h t E S D I O N 1 8 - 1 6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000241_0001
u n S d DGEFK T GYQV EQE F T N LP LK MS LTQTYK C E d i c TA n A L KAA G RYNKD T CRH K P T EQP L S AP F S VGK NYLLQ LRY P LAAV QYP T D LLAGLGVS HAGR KWVHLANT I I K RG S L T a o D d n i P RQRI Q i RQAS EHE AKI VDP LKQL TYTPKGLGR VAS S P LLLS AGK GT I WS L I QP P P F LQGKAI QKAGP E K AGA WLANG c m a A T G GE S DK QTT PA AVPGVGQ HLGYGCA QTQP S S GI TQL L LPAE K VK S H I K P F S D R L P P L D LYT EKDAG PA RKP MQDS KP F L I N TKKG RQI K AG ERKPDL QEY RG AV L TQQDYWAR LAS P o A LAP NTI R DTLYL HRHTVGA FKFYLKV RL C W I MR YF L MNK n i E S S LLLS LA S ALDNLT Q I GQMGQLAVV L DNTVKI ARP S R A me GAKW a r L I RAN I L K LR N T Y P R LWRQK PQE F V P A I A THHE LE NS G S f DN TKK A GG RK PQ DL R W T PI Q I I HG P N P C F T LF DQQV CC I T I E F L D P P KR RA AE VD L Q LDT AE TAGK QH I L L G RN E G P o u s g o l ni c t si n oi s t 6 3 i d t L L t p i a nr c B 1 s D- 2 4 . 3 : 2 e s e D C 0 0 0 e l e r - 8 b p 2 a e Q : 8 6 T h t E S D I O N 2 8 - 1 6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000242_0001
u n S d R A QLHLN S S RLE DEAA WG LQP LML LR RKI Q d i c G n A G S TDAQL I N FAGI VAP DVAQKS S S V LP AT A EL A LKTR VDNL RRL G YQF P ATWP RT DQQS EH KP LGQAA P P GAS DKG a o d n i G PQAAQHN i R E H PQT I HS TAGQNKT YRF L GG RT A NP S NMS WN WL TVAGKNR TDLE R L LLGAMS LLTETT YGP GQGVK C c m K RMPQF I D a A R GR RDPQP S QT NS K LEGL TGP P S VWEPDNK YQEQF S VGL L Q P LK RYAAVT AHLA YP L WVA L S P N S L TH I L I S o Q RAQ DKH DS D S TQCRHS A FNY QLYP L PQGDKR GV ES LK L n i AG S S AS P PAGL D EKP LP LKL T G TAKI QLG APAGA W KKG R e L S I T G P L S M RNF S V VNNY S AI VD QP P F LQ YI T PNWS L I VP QGI L L LL KA YP TM E QKKA EKS DP G F L KQQI NAG DT ERK RP m a r K AK RS EHQD A H VN S FKKAV NV I T I QP TG QPG RS S L TQ P P DAKP EYDL TKVLYW I ARL F f u s L P R V Q A A V Y L V R C W V M Y L o g o l ni c t si n oi l s t l u f i d 1 L L t p i t : nr c s a B 1 D2 4 . 3 0 2 e s e D - N 0 0 e l e r - 8 b p 2 a e Q : 8 6 T h t E S D I O N 3 8 - 1 6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000243_0001
u n S d i T d c E n A T I a o A d n i A ic m K a A R A o A n N i QI DL me a r AR S fo u s M N g o l ni c t si n oi s t i d L t p i r c 1 . : n 2 e s 3 s e 0 0 el e D r 0-8 b p 2 8 a e Q D : O 6- T h t E S I N 6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000244_0001
u n S d G QV EQE F T LP LKAMS N LTQTYK C EDHT V S NI d i c TY n A EQC P R LH P S A F S VN GKQYL AAVT NYLL LR P L YP LLAGLGVS HAGRQS YRNMF KWVHLANT I I KG S L T Q S DQVNP R a o AKI d n VDP L i WS L I QP i P P KQ F L T LQY GTPQ KAKGD I QLGR KAGVAS P P ES LL KLS R GKS P S LQHVQ GA WLA ANGGH S QGNL c m AV a A QP P TGV QP GQGI S S L LPAE K TQL DLYT EKDAGA P MQDS KP F L I N TKKG RQI KDT P AG P L ERKPDL GTLL L GVY QS NN o F S D R L P P DTLYL H LPA RHRK TVQEY GA FAV L TQQDY KFYLKV RL C WW I AR LAR S P E KE MR YF MNKGR S L T S QNS Y QGN LP ni S ALDNLT Q I GQMGQLAVV L DNTVKI L ARP S R GAAQT S NG me L K a r W LR N P Q I T Y G P R N L CW TR FQK QQVP T I FQE DF V P P A R I AA ET DHHE QDT ELE N AGK LG S G N S GTDAQS DVAQLHM N L F f u s T I I H P P F L D C C I E L P K R A V L L A T QH I L R E P Q A A QL I S S o g o l ni c t si n oi l s t l u f i d 1 L L t p i t : nr c B 1 s a D2 4 . 3 0 2 e s e D - C 0 0 e l e r - 8 b p 2 a e Q : 8 6 T h t E S D I O N 4 8 - 1 6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000245_0001
u n S d i I c H dn A A a o A d n i P i S c m P a A Q T o R n i P Q me N a r S f u s A I o g o l n n i c t si o si di t L t p i r c 1 . 3 : n 2 e s s e 0 0 0 e l e D r - 8 b p 2 8 a e Q D : O 6- T h t E S I N 6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000246_0001
u n S d RGAGQ RLLD LMLAS LKLGKCI HPDGEF T RDNGL P d i c n A G S G I DS DVGVRLE D AAAP V L EAAAWG TLQPVTL TAAAKG VP ELRRL G P AT P L N D a o P ML TL K RRY K GP K E LGY QP T S D E RTQ d n i G PQA i R EVLKT HGQNRT D RN F L LYQF K P S KY A L VWQR AD PQQRI HE G LLEK WLDGG R RTTGGGA KP RGAS DE KGT I P RAAI S c m K a A RRMPM GR RDP LWPNWE TDLE KL L o Q AGP S V P LKAM EQF P S VN S N TE S TTGPA K RRDWV GLQYLAAT LL AGQGYK CA GR RVA QP n i AGR L S S S TGYQ I TT EQC P R LHS ANY P F L LR LA LKQTYP TPQYV P L WVHL LAVS H I KGDKR VA E S P NT L I K Q S R AGA S E L P T F S T T GP P K NAI VDP P F L LQG L KA AI QLGGPAS LL LANLGDDS A me KP RHWS L I QPQGI L L LYP E QK EKA DA P G F LGAKWRAI TP RLK a r AK R S E A AV P GV P G S S T Q P D A KT P M E Y DK S I NKKGKQLKKS L f u s K Q Q D T A G R P D R A R S W P o g o l ni c t si n - - oi s t L T i d - 0 L 4 T - 1 L t p i : nr 1 t 0 L B 1 t 6 3 L B 1 . 2 e c s a D s e D I 2 - Q 4 a DP - X2 4 3 0 e l e D r N N 0 0-8 b p 2 a e : 8 Q 6 T h t E S D I O N 5 8 6 - 1 8 1 6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000247_0001
u n S W d PNWWTDLEL LLKAMS N T TY CA KA S P PKTA d i c S V n A QE RQE HF PVNKQ P AS YLAVT LLGQ LGVK S HA R ENGAK I V PD NYGLL LR LAYP L WAHLANT I I K G S E N DI GQGP H I K a o C n L S F i PDP KQTYP TPQGD P L F L LQGKAKI QLK V GR KAGVAS P P ES LL KLS R GVN S R LL L D L GA WLA AN P R RGS VR P E D d i c m V a A L I Q VP PGI GP GQ S QL L LPAE K D LYT EKDAGA P MQDS KP F L I N TKKG R AG ERKPQI K RQ DL L P AAT V L R R R S S LK G NR V T o QRS TP YL LP LPA RHRK TVQEY GA FAV L TQQD K KFYL V RLY CWW I AR MRLAS GRGT Q S WKY YF L MN Q AG S L I N LM GLPN ni D LDNHL T Q I GQMGQLAVV L DNTVKI ARP LGCGGP S QV EW me LNY a r R RWRQKVP I QE F V P A I A THHE TLE GKNS S TP E RTYCRQ H QT GP LT F QQ CTF D P RAE DQDEAHLGN KKNS LEQ KP L S f N C L D C I E L P K R A V L L A T Q I L R E A R I A A I V D P o u s g o l ni c n -L t si oi si d t p T i x -1 8 5 5 L L t : nr c T t s R a 2 e s e D P B 1 . X2 4 3 0 D - N 0 0 e l e r - 8 b p 2 a e Q : 8 6 T h t E S D I O N 7 8 - 1 6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000248_0001
u n S d GL LKA F NP G R QP L P A TES T R S I NTWVL I QG S QD K d i c S I D E S KEGF LL LAYL A QGI n A GRT ATALGKC G I H a o T P LD ADHS GNS S APGVP S TP AL EFKGT S GVRQP TQR L P LP R TV HRHGQ d n i MH i E LAS LK AAAWG T L G LQP P A P T LPV LTKAANKD P GDS GP F S T DYL L T Q I QK WL P M RTLRRYKQT P RNS TDALDNRWR F QQ c m P a A D EL R NL RR F F LY KAVGQAADQQR GQ GRTTGAGKP NP RGAS I D EHE K KGT I RRR S S GKS S L K L L RNY P LT C C Q T N C F LD ADI Q o W DLE R n i EL F TD P NKL L Q P LKAMS LTE S YL TTGP CA QRS K N W T P II I G P P FQL LA AAVT LAGQ LGYKHA AG S VNLGLHKVAGS D E S K S VGL LR LQYP L WVHLAVS T I I K LG TS S P P KKT PDI GRTK me ANYL TYP TP G AKQD LK GR GV P AS P N LS R TP S NGAI VHK HAA a r F L KQ L Q G K AI E K K A A GE AS GL AL K WLA KK TL GQGP I L M L A S f u s R A N A R V A S R L L D D E A L A o g o l ni c t si n - - oi s t L T L i d - 3 T - 3 L t p i x : nr c T 1 t 8 6 L B 1 t 8 6 L B 1 s Ra 2 e s e D P X2 4 a DP X2 4 . 3 0 D - e N - N 0 0 l e r - 8 b p 2 a e Q : 8 6 T h t E S D I O N 8 8 - 1 6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000249_0001
ob m e g n iti mil-n o 8 n 4 f 2 o se Q c E S D : I O N ne u L q F LWP D KGT I es RPPVGQRA P PQAD d RK TGPA i LT A L TG LKAGANRGS A K TYK CA MS LT GEGVS H I K c Q P LAVT LA Q HLANTI R a e RYAYP L WVA L P LLS LA c c i n P L e e TPQG K DK l u KAI QLGR V GP E S LKW AS AKG RA KQNI c q u e PAE KKA S T P MQEKDAG P F L I G NK AG ER RP LDL AR YDS KQQDT WAR F M S n d QEAV L TKVLYI MYI LR N d i c n A GA FKFYL V RC WV a o QL d n MG EAVL D TNTE K LEA P GKNS S i VP i I QF P V P A I A H ED HT LQDE TA QH I L L G RN E c m T F DKR A E L P T RA AV AL L LGA DT VTHAAI LA L I EL a K RGL o R C WV FWDP MH PAE QE P P LAA TE A AI ADS L T HS n i A E TGNP G RAYLPGL FKGT AF LL LG I HPQAE AANKD me a r LKL W GT G GLKC TLD AQ T P P V L ML T TK RY RKI QP T f u L R Q S E H E o s g o l ni c t si n oi si d t p i L t nr c 1 s . 3 : 2 e s e 0 D 0 0 e l e r - 8 b p 2 a e Q : 8 T h t E S D I O 6- N 6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000250_0001
o s g c c u c c c c c u c u g u a L R t a n c c g u c a u u a c a a u c u g u u u c g t a c A C H s E I GL e g g g g u u c c c c c u c c a g c c c c c g t t a a g a a a a u c c c c a g u c a a a c c g g g a t A i H C H LQC R D F F mi a g c c c a g c g u g g g a c c c c c g d s e a c c a g u g t/ g a u g a g c a u g a c a a u g a g g G s A i H LA a M t H E L RD o c b n g e g g g a c c u u g u u a g g u u a c c a g c g u c a g g g c g a g u g a c a c g g g g A g G F a c T T si D K H E I Q I L m u e q e a c g g a g u a c c g c c g a a c g g -6 g G s S g a g c g u g u c u c u c u c u a g c g c a a u c c g g 5 a G i N L H P D L g a c u c u a / a T H T K V niti n m o e i i l t H e t 6 5 g - a t p y - p n i r D ) a o c s P Al H p Dr H e 1 o D r s P e i o N e n T V d L ANme P Am i r M AOS gi l Am i r H i d i li M 1 . . D GR ( T G P F T O G P x 5 m a c a W M 3 0 5 0 0 e - l D b I 8 2 a Q: T E SO 1 N 0 2 2 0 3 4 5 6 8 6- 2 0 2 0 2 0 2 0 2 6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000251_0001
L T E GKD HT YQ AP KDV L A T D V ED L A T D V E AP W T A D P ADGS RYKTAKI AYQS L L R TK AYQS L L R T S LYP HT T WN T I YL P S DAE WD NGA P KD T VN WD NGA P KD T T F AEGQDA E A AR LQKAF T K VDAI QGG P MG LAL L Y A LGK AT S LAGN AP F KE P N AAKK S I Q L P F G I RYKTAK AT S I RYKTV A P GYL G P S L D I QGGYL P S L D AYLAE GK P F P GAYLA L L Q E L P YADKDGMV P QR T R DKM E LAAL L DKM E LAAL LG L RDA E L R S K S V R L E A RYQL S S KP N V AG AK NS KP N V A A DP V DT F VP QD R L P T F R H GWN LVDGM E P QR K S I LVDGM E P QK R A TAWLWQ L G T E GP S WP R L T ARYQT L WP R L T ARYQ VKP L D TYAT L P E F L P W P MQ L R S T L WQ F R H G L R S T L W F R Y T E F H P L V Q M T A AI F G R L R N ML P E L G W T E W N ML P E Q L G WH T G E la D c I l D it Q 7 V t a I L n a ) c it Q 1 V t L n a V t L n n E e d S : I o t O M N MT i r Ra B 2 n E V4 e ( d S : I o t O M N MT i r a Ra M V MT i r Ra 1 V . 3 0 0 0-8 2 70 8 9 8 6- 2 0 2 0 2 6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000252_0001
K DQ L TK S I R DQ L TK DQ N WNGA P KDVN WN T YL P S DAE WNGA P KDVN WNGA KI AT S I RYKT K A GGAL LGK* AT S I RYKT K AT S I R E Q K P F G P GYL S MG P L DA AYLAI E QG GK P P ML YL LAGNQ F KE P NAAAK K S I L P F G YL S P G DAI Q MG P L AYLAE GK P F G Y P GG G DK L L KNS KENAALAG* DKDGMV L E P QR QT R S DK L L KML KENAALAG* DKEN S I P L LVDGM E V P A QK R KN S I S V R T L L P A RY QF R H G L S P EWN LVDGM E V P A QK R KNS P S I L LVDG T L R S W T P L R L T A R F RYQT L R S W T LW L E F L L G P W T P MQGP WP R L R S S T L T A R F RYQT L R S W T P L R L T W N ML P W E Q L G WH T G E W N MP T A AI F G R L L R N ML P W E Q L G WH T G E W N ML P W E Q L V t t t t L n a i V L n a i V L n a V L n a M r M r i r i r MT Ra V MT Ra M V MT Ra M V MT Ra 1 V . 3 0 0 0-8 2 0 8 1 1 6 2 1 2 3 - 2 1 2 1 2 6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000253_0001
L P KR D TK E L I R L E S LAA W TH L A P E I R L E S LAA W TH L A P E I R E S LAA W THD YKTVN HGC T F YP HT T HGC T F YP T T HGL T F YP A T P L S DAKI LQF ARGQ E LQF AR H GQ E LQC F AR H GQ T P L LAE R D F R KDA F T I R D F R KDA F T I R D F R KDA F E AYLGK H LAA L LQ A AH LAA L LQ A AH LAA L LQ A T I AAL * E L L QK V L DAAE L L Q KV L DAAE L L Q KV L DAA MV AG P AKKNDRDL R E S I E QKDP DA P Y E L R R DRDL R E P Y R E QKDP DA E L R R D E RD Q K L R P E P Y DA L RA E RQR T L I I L DT VVP S I I L T V P S R I I LDT V E P S K A RYQL R NHD KF QNANHDDKF VQDANH DDKF VQN R F H G S L P LVP LAD TAL P LVP LAD TAL P LVP LA TA G W T E W N T K V Y T E H T Y Q T K V Y T E H T Y Q T K V Y T E HD T Y A V t V t t L n a L n V L n M MT i r a i Ra M r a i V MT Ra M r V MT Ra 1 V . 3 0 0 0-8 2 41 5 6 8 6- 2 1 2 1 2 6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000254_0001
E L I R E WH L R E WH L R E WH L R E K HGL C S L T F AA YP T HA P E I T T E HGL C S L T F AA YP T HA P E I T T HGL C S L T F AA YP T HA D E I T P HGL S L T F A LQF ARGQ A T LQF ARGQ A E T LQF ARGQ A T LQC F ARY R D F AR D QK F I R D F A R KDF I R D F A R KDF E T R D F A RG HLAL L VAAAH LAL LQVAAAH LAL LQVAAI H LAL LQ E LDL Q K L DYAE LDL QK L DYAE LDL QK L D AE LDL QK DR I I K L R E A P L R DR K L R E A P L R DR K L R E A P YADR L R E E Q LDP DE P R S R E Q LDP DE R R E Q DP DE L R K E QKDP D N L H DT VV P D L Q AI I DT VVP S Q AI I L DT VVP S Q R I I L DT V VKF A D TANH P D LVKF A N TANHD KF N NHD KF T K V YP T L E HD T Y QL T K V YP T L E HD T Y QL P LV T K V YP T LA E HD TA T Y AL P LV T K V YP T L E V t n V t n V t V t L a M i r L a i a M r L n a i V M Ra M r L n a i V M Ra M r 1 MT R T T V MT Ra V . 3 0 0 0-8 2 7 8 1 8 2 1 9 0 6- 2 1 2 2 2 6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000255_0001
AW WD K T D K T W P T N I L E N QHA D T P T ANGA P T S D I RYKTVN WNGA P DVN TGYP S AK AT S I RYKTAK AG GAL D LA W LGK AT G KDA VAF E Q T P F G P GYL S G P L D LAI E QP G P GYL S D I Q GP L LAE P P ML Y KENAALAG KKNI QP L P K L DAI P KMLAYLGK F KMLAYLGK F YAD KP GMVAR S F K KENAALAG DKENAALAG D D L E P QQT R S D A E L AS P P R S R LVDGM E V P A QK R KN S I S LVP DGM E V P A QK R KN S I S L V P R T A F R RY H G L S V WN L P VQ A N R W T P L R L T A R F RYQT L L R S W T P L R L T A R F RYQT L L R S W T LWQ L E F L L G P W T E P P QGS S W T L L HD TA T Y A ML P W E Q L G WH T G E W N ML P W E Q L G WH T G E W N MP T A AI F GM R L L R R N MP T V t L n a V t L n a V t L n V L M i r i r a i r MT Ra M V MT Ra M V MT Ra M V MT R 1 . 3 0 0 0-8 2 1 8 2 2 6 2 2 3 4 - 2 2 2 2 2 6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000256_0001
S I R YL DQ L R K DQ L R K Q L R P S DAE W GA P KD TN W GA P KD TN WDGA P KD T GG L LGK AN T S RYKTVK AN S RYKTVK AN S RYKTV MLAYLAG Q I YL S DAI QT I YL S DAI QT I YL S DA EP NAAL K KNI P GGG P L LAE P GG P L LAE P GG P L LA DGM R L E VA P QR S L F P MLAYLGK F P MGAYLGK F P MGAYLG T A QT R DKE AALAG DKE LAALAG DKE LAALA F R RY H G L S S KP N VA WN LVDGM P QK R KNS KP N S I LVDGMV P A QK R KNS KP N S I LVDGMV P AK W EQ F L L G P W T E P MQGP L R S S W T P L R L T E A R RYQT L L R W T P L R L T E A R RYQT L L R W T P L R L T E A RQR RYQ M P W E Q F L G WH T G S E W N M P W E Q F L G WH T G S F A AI F G R L R N L L E W N ML P W E Q L G WH T G E tn a i V t t t r L n a i V L n a i V L n a a M V MT r Ra M V MT r Ra M V MT i r Ra 1 V . 3 0 0 0-8 2 52 6 7 8 6- 2 2 2 2 2 6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000257_0001
K DQ L TK DQ L TK S R S R N WNGA P KD N WNGA P KD N WN T I YL P S DAE WN T I Y KI AT S I RYKTVK AT S I RYKTVK A GGAL L E Q LGK A GG K P F G P GYL S DAI Q G P L AYLAE P G YL S P G DAI QG G P L LAE P P ML Y KENAALAG K KNI QG P ML L P KEN G DKML LGK F KMLAYLGK F KNS KP GMVAR S KENAALAG DS KENAALAG DS DL E P QQT R F S DKP DG L S I P T L LVDGM E V P A QK R KN S I P L LVDGM E V P A QK R KN S I V R T L L P A RY L S V R T QF R HG EWN L P Q L R S W T P L R L T A R F RYQT L R S W T P L R L T A R F RYQT L R S W T LW L E F L L G P W T P QGP S S W T LW L E F L L W N ML P W E Q L G WH T G E W N ML P W E Q L G WH T G E W N MP T A AI F GM R L L R R N MP T A A V t t t t L n a i V L n a i V L n a V L n a M r M r i r i r MT Ra V MT Ra M V MT Ra M V MT Ra 1 V . 3 0 0 0-8 2 8 8 2 9 6 2 2 0 1 - 2 3 2 3 2 6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000258_0001
L DQ L R K DGA P K Q L R K P S AL D LAE W Y GK ANGA P K T S RYKD T TVN W S RYKTAI WDGA P KD TN K AN T I YL S DAE AN S RYKTVK A L ALAG MVAK R KN S I Q L P F G I P GYL S DAI Q G MG P L AYLA L E GK P F G P MG P L LGK QT I YL S E LAYLAG P GG DAI AAL K KNI F P MG P L AYLA L E GK E P Q T R DKE LAALAG DKP N VAR S L DKE LAALAG ARYQL S S KP N VAK KNS KDGM P Q T R S KP N VAK N F R H G N LVDGM P S I LVR L E RYQL S LVDGM P K S I G P W T EW QGP S W T P L R L T E ARQR RYQT L L R W T P T LWQA F R H G T EWN P WP L R L T E A RQR RYQT L L R I P F GM R L R S L WQ F H G S L E L G W Q GS T L WQ F H G S L R N M P E L G W T E W N M P F L P P M L R S M P E L G W T E W N V t n V t V t L a L n a L n a M MT i r Ra M V MT i r i Ra M r V MT Ra 1 V . 3 0 0 0-8 2 23 3 4 8 6- 2 3 2 3 2 6 7 8 4 5 1 C 2 P 0 - 1 8 0 8 5 4 5 1 6 3 1- 1: 0 . f 0 e 1: R. f y e e l R o s F c i mo n e G X0 1
Figure imgf000259_0001
KI L T S L A LGC NT K K I LDT L A TK GLAH VNL DW RLA L S DHAQ VNA L WLACQ NTG P Y A Q I H P Y LD LK T P RQQK L T P D RA H S T N P H GKT L LGEG P S N R P QQK T Q A I Q H AVL A E I EVE A C T H VL G EKI L L GL EG KT EK L P LQA I E R P QH I A KT AKEVE AP C PII H I N F LYALAI S L P LYAAP L D T G LKVL S L P I H E LQA I E I N L T F LY P R H 8 Y ALQI I 5 2 GR I K L D L GA KAP AS S P E P Q A M T P E T L VL L V S YQ AP V P QML T R QRNKD R L A P P L L F A QR S Y KQR TGK RKKV VNF QP GC NAL RKN KVD P L A P A P P L AVGK P L WP AKA K VN V F QGC P N L AF L L E RQP Q A C NWV L A L EGK QQWP L AKA GL L I F LGV P W I L R NW A GL P C F GVAK MDRQ A GQTATAG V E L MDL R I QLAV W I L L GVWKG P V F T I E GQ G TAT GAP G VA T E P R TK L AK P Q D V E TD TV P WKP V P F E L QT I AYQS T K L L R T P KDVN E R T AYLK A K E S T D L V E R TD WDGA TAK WDQ L TK AN S I RYKS DAI E ANGA P K RYKD TVN QT PF GGYL P MG P L LGK QT S LAYL LAGN G I GYL S L DAKI DKE S P NAA I P F P G P AYLAE GK VAK R K S L DKM E L L NAALAG LK VDG L M E P QQT R S LKP GMVAKKNI WP R T A RY L S V R G N P D R L E P RQR S T L T LWQ L F H T EW W QG P T L T A F RYQL R S ML P E F L G P W P M L R S S ML P W E Q L G WH T G E W N V t t L n a V L n MT i r a M i Ra M V MT r Ra 1 V . 3 0 0 0-8 2 53 6 8 6- 2 3 2 6 7 8 4

Claims

Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    WHAT IS CLAIMED IS: 1. An engineered reverse transcriptase (RT) polypeptide comprising: (a) an RT polypeptide; (b) a DNA binding domain, wherein the DNA binding domain is from a molecule capable of binding a minor groove of a nucleic acid; and (c) a linker connecting the RT polypeptide and the DNA binding domain. 2. The engineered RT polypeptide of claim 1, wherein the DNA binding domain is located at the N-terminus of the RT polypeptide. 3. The engineered RT polypeptide of claim 1, wherein the DNA binding domain is located at the C-terminus of the RT polypeptide sequence. 4. The engineered RT polypeptide of any one of claims 1-3, wherein the linker is: (a) a glycine-serine linker selected from the group consisting of (GS)n, (GSGGS)n, (SGGSG)n, (GGGS)n, (GGSG)n, (GGSGG)n, (GSGSG)n, (GSGGG)n, GGGSG)n, and (GSSSG)n, and wherein n represents an integer of at least 1; or (b) GGGS; or (c) SGGSG. 5. The engineered RT polypeptide of any one of claims 1-4, wherein the DNA binding domain specifically recognizes adenine-thymine-rich region on a nucleic acid molecule. 6. The engineered RT polypeptide of claim 5, wherein the DNA binding domain specifically recognizes oligo(dA) or oligo(dT) tracts on a nucleic acid molecule. 7. The engineered RT polypeptide of any one of claims 1-6, wherein the DNA binding domain comprises at least one AT-rich interaction domain. 8. The engineered RT polypeptide of claim 7, wherein the DNA binding domain comprises at least 2, at least 3, at least 4, at least 5, at least 6, at least 7, at least 8, at least 9, or at least 10 AT-rich interaction domains. 259 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    9. The engineered RT polypeptide of 7 or 8, wherein the AT-rich interaction domain comprises a core sequence, wherein the core sequence is a two base core sequence, a three base core sequence, a four base core sequence, or a five base core sequence. 10. The engineered RT polypeptide of claim 9, wherein at least one of the bases of the core sequence comprises an arginine; a glycine and an arginine; a proline and an arginine; a lysine and an arginine; or any combination thereof. 11. The engineered RT polypeptide of any one of claims 7-10, wherein the AT-rich interaction domain comprises a GRKPG (Gly-Arg-Lys-Pro-Gly) repeat, a RKRGRPKK repeat, a KKRGRPKK repeat, a RKRGR repeat, a GR*R/PPK repeat, a GR*RPK repeat, a GR*PPK repeat, a KRPR* repeat, or a K/RKRGRPKK repeat. 12. The engineered RT polypeptide of any one of claims 7-11, wherein the AT-rich interaction domain comprises a core sequence comprising an amino acid selected from the group consisting of SEQ ID NO: 11-24. 13. The engineered RT polypeptide of any one of claims 1-12, wherein the DNA binding domain is a DNA binding domain of any one of Saccharomyces cerevisiae datin (DAT1), high mobility group AT hook 1 (HMGA1), lysine-specific methyltransferase 2a ( KMT2A), Myocyte Enhancer Factor 2C (MEF2C), Heterogeneous Nuclear Ribonucleoprotein D (HNRNPD), Structural Maintenance of Chromosomes 1A (SMC1), Structural Maintenance Of Chromosomes 2 (SMC2), Caenorhabditis elegans tbp-1, Drosophila melanogaster D1 protein, Salmonella typhimurium Hin recombinase, S. typhimurium Gin recombinase, S. typhimurium Pin recombinase, or S. typhimurium Cin recombinase, or a combination thereof. 14. The engineered RT polypeptide of any one of claims 1-13, wherein the DNA binding domain is from a S. cerevisiae DAT1. 15. The engineered RT polypeptide of any one of claims 1-14, wherein the amino acid sequence of the DNA binding domain comprises a DNA binding domain consensus motif set forth in SEQ ID NO: 13, 14, 16, or 22. 16. The engineered RT polypeptide of any one of claims 3-15, wherein the DNA binding domain comprises: 260 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    (a) a full-length DAT1 sequence or SEQ ID NO: 2; (b) an N-terminal truncated variant of DAT1 (D90) comprising the first 90 amino acids of the full length DAT1, or SEQ ID NO: 3; (c) a truncated variant of DAT1 (D60) comprising the first 60 amino acids of the full length DAT1 or SEQ ID NO: 5; (d) a truncated variant of DAT1 (D48) comprising the first 48 amino acids of the full length DAT1 or SEQ ID NO: 6; (e) a truncated variant of DAT1(D36) comprising the first 36 amino acids of the full length DAT1 or SEQ ID NO: 8; (f) a truncated variant of DAT1 (D35) comprising the first 35 amino acids of full length DAT1 or SEQ ID NO: 9; (g) an amino acid sequence having at least about 90%, at least about 91%, at least about 92%, at least about 93%, at least about 94%, at least about 95%, at least about 96%, at least about 97%, at least about 98%, or at least about 99% sequence identity to any one of SEQ ID NO: 2, 3, 5, 6, 8, or 9; or (h) an amino acid sequence having 90%, 91%, 92%, 93%, 94%, 95%, 96%, 97%, 98%, or 99% sequence identity to any one of SEQ ID NO: 2, 3, 5, 6, 8, or 9. 17. The engineered RT polypeptide of claim 16, wherein the DNA binding domain comprises the amino acid sequence of SEQ ID NO: 11 (GRKPG); optionally wherein the DNA binding domain comprise at least 2 domains or at least 3 domains comprising SEQ ID NO: 11. 18. The engineered RT of claim 16, wherein the DNA binding domain comprises a mutation in any of one of SEQ ID NO: 2, 3, 5, 6, 8, 9, or 11, and wherein the mutation is selected from a substitution, an insertion, a deletion, or any combination thereof. 19. The engineered RT polypeptide of claim 16, wherein the DNA binding domain comprises SEQ ID NO: 2. 20. The engineered RT polypeptide of claim 16, wherein the DNA binding domain comprises the amino acid sequence of SEQ ID NO: 3, 8, or 9. 261 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    21. The engineered RT polypeptide of any one of claims 1-20, wherein the RT polypeptide sequence comprises the amino acid sequence of SEQ ID NO: 7, and further comprises a combination of mutations selected from the group consisting of: (i) E69K, L139P, E302R, T306K, W313F, T330P, and N454K; and additionally one or more of M39V, P47L, Q91R, M66L, F155Y, D200N, D200E, H204R, G429S, L435G, L435K, P448A, D449G, H503V, D524N, T542D, E545G, D583N, H594Q, L603W, L603F, E607K, E607G, P627S, H634Y, H638G, A644V, D653H, K658R and L671P; and (ii) E69K, L139P, D200N, E302R, T306K, W313F, T330P, L435G, P448A, D449G, N454K, D524N, L603W, and E607K; and additionally one or more of M39V, P47L, M66L, Q91R, F155Y, H204R, G429S, H503V, T542D, E545G, D583N, H594Q, P627S, H634Y, H638G, A644V, D653H, K658R and L671P. 22. The engineered RT polypeptide of any one of claims 1-21, wherein the amino acid sequence of the RT polypeptide sequence is: (a) at least 90% identical to SEQ ID NO: 1 or 143; (b) about 90% to about 99.99% identical to SEQ ID NO: 1 or 143, about 92% to about 99.99% identical to SEQ ID NO: 1 or 143, about 93% to about 99.99% identical to SEQ ID NO: 1 or 143, about 94% to about 99.99% identical to SEQ ID NO: 1 or 143, about 95% to about 99.99% identical to SEQ ID NO: 1 or 143, about 96% to about 99.99% identical to SEQ ID NO: 1 or 143, about 97% to about 99.99% identical to SEQ ID NO: 1 or 143, or about 98% to about 99.99% identical to SEQ ID NO: 1 or 143; or (c) about 90%, about 91%, about 92%, about 93%, about 94%, about 95%, about 96%, about 97%, about 98%, about 99% or about 99.5% identical to SEQ ID NO: 1 or 143. 23. The engineered RT polypeptide of claim of any one of claims 1-22, wherein the RT polypeptide sequence comprises: (a) an amino acid sequence that is at least 95% identical to SEQ ID NO:1, 7, or 179, and; (b) a combination of mutations indexed to SEQ ID NO:7 or 178 selected from the group consisting of: 262 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    (i) a combination of variants consisting of a T542D mutation, a D583N mutation, an E607G mutation, an A644V mutation, a D653H mutation, and a K658R mutation; and (ii) a combination of variants consisting of an E545G mutation, a D583N mutation, an H594Q mutation, an L603F mutation, and a S679P mutation. 24. The engineered RT polypeptide of any one of claims 1-23, wherein the amino acid sequence of the RT polypeptide sequence comprises E69K, L139P, D200N, E302R, T306K, W313F, T330P, N454K, H503V, D524N, L603W, E607K, and H634Y. 25. The engineered RT polypeptide of any one of claims 1-24, wherein the amino acid variation(s) are at any one position or combination thereof as identified in an alignment of SEQ ID NO: 1 or 143 to any one of the RT polypeptide sequences in Table 1 or Table 2. 26. The engineered RT of any one of claims 1-25, wherein the RT polypeptide sequence comprises: (a) M39V, M66I, Q91R, I347V, and H594Q substitution in SEQ ID NO: 143; or (b) SEQ ID NO: 129 (SOLD 034). 27. The engineered RT polypeptide of any one of claims 1-26, wherein the RT polypeptide sequence comprises M39V, T542D, D583N, E607G, A644V, D653H, K658R, and L671P in SEQ ID NO: 143. 28. The engineered RT polypeptide of any one of claims 1-27, wherein the RT polypeptide sequence comprises: (a) M39V, T542D, D583N, E607G, A644V, D653H, K658R, L671P in SEQ ID NO: 143; or (b) SEQ ID NO: 111 (SOLD 025). 29. The engineered RT polypeptide of any one of claims 1-28, wherein the RT polypeptide sequence comprises T542D, D583N, E607G, A644V, D653H, K658R, E545G, D583N, H594Q, and a L603F in SEQ ID NO: 143. 30. An engineered RT polypeptide comprising: (a) an amino acid sequence that is at least 90%, at least 92%, at least 95%, at least 263 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    97%, at least 98%, or at least 99% identical to: (i) an amino acid sequence of an RT disclosed in Table 1, or Table 2; or (ii) SEQ ID NOs: 27-61, 63, 65, 67, 69, 71, 73, 75, 77, 79, 81, 83, 85, 87, 89, 91, 93, 95, 97, 99, 101, 103, 105, 107, 109, 111, 113, 115, 117, 119, 121, 123, 125, 127, 129, 131, 133, 141, 143, 145, 147, 149, 151, 157, 159, 172, or 173; and (b) a DNA binding domain comprising an amino acid selected from the group consisting of SEQ ID NO: 2, 3, 5, 6, 8, and 9. 31. An engineered RT polypeptide comprising: (a) an amino acid sequence of an RT disclosed in Table 1 or Table 2; and (b) an amino acid sequence of DNA binding domain disclosed in Table 1. 32. The engineered RT polypeptide of any one of claims 1-31, wherein the engineered RT polypeptide comprises: (a) the amino acid sequence of any one of SEQ ID NO: 174-188; (b) an amino acid sequence having at least 90%, at least 91%, at least 92%, at least 93%, at least 94%, at least 95%, at least 96%, at least 97%, at least 98%, or at least 99% sequence identity to any one of SEQ ID NO: SEQ ID NO: 174-188; or (c) an amino acid sequence having 90%, 91%, 92%, 93%, 94%, 95%, 96%, 97%, 98%, or 99% sequence identity to any one of SEQ ID NO: SEQ ID NO: 174-188. 33. The engineered RT polypeptide of any one of claims 1-32, wherein the engineered RT comprises an amino acid sequence that is at least about 90% identical to an amino acid sequence selected from the group consisting of SEQ ID NO: 27-61, 63, 65, 67, 69, 71, 73, 75, 77, 79, 81, 83, 85, 87, 89, 91, 93, 95, 97, 99, 101, 103, 105, 107, 109, 111, 113, 115, 117, 119, 121, 123, 125, 127, 129, 131, 133, 141, 143, 145, 147, 149, 151, 157, 159, 172, and 173. 34. The engineered RT polypeptide of any one of claims 1-33, wherein the RT polypeptide is 42B L (SEQ ID NO: 145), 50A+G (SEQ ID NO: 147), SOLD 022 (SEQ ID NO: 105), SOLD 023 (SEQ ID NO: 107), SOLD 025 (SEQ ID NO: 111), SOLD 031 (SEQ ID NO: 123), SOLD 033 (SEQ ID NO: 127), SOLD 034 (SEQ ID NO: 129), SOLD 035 (SEQ ID NO: 131), SOLD 001 (SEQ ID NO: 65), and SOLD 33 VDG (SEQ ID NO: 173), or an RT polypeptide set forth in SEQ ID NO: 143, or SEQ ID NO: 172. 264 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    35. The engineered RT polypeptide of any one of claims 1-34, wherein the engineered RT comprises at least two DNA binding domains. 36. The engineered RT polypeptide of claim 35, wherein at least one DNA binding domain is located at the N-terminus of the engineered RT and at least one DNA binding domain is located at the C-terminus of the engineered RT. 37. The engineered RT polypeptide of claim 36, wherein the at least two DNA binding domains are both located at the C-terminus or N-terminus of the engineered RT. 38. A recombinant reverse transcriptase (RT) protein comprising a RT polypeptide, fused to a DNA binding domain, wherein: (a) the RT polypeptide and the DNA binding domain are separated by an amino acid linker, and (b) the DNA binding domain is fused to the C-terminus of the RT polypeptide. 39. A recombinant reverse transcriptase (RT) protein comprising a RT polypeptide, fused to a DNA binding domain, wherein: (a) the RT polypeptide and the DNA binding domain are separated by an amino acid linker, and (b) the DNA binding domain is fused to the N-terminus of the RT polypeptide. 40. The recombinant RT protein of claim 38 or 39, wherein the RT polypeptide is any one of the RT polypeptides listed in Table 1 or Table 2. 41. The recombinant RT protein of any one of claims 38-40, wherein the DNA binding domain is a DNA binding protein selected from the group consisting of S. cerevisiae datin (DAT1); high mobility group AT hook 1 (HMGA1), lysine-specific methyltransferase 2a ( KMT2A), Myocyte Enhancer Factor 2C (MEF2C), Heterogeneous Nuclear Ribonucleoprotein D (HNRNPD), Structural Maintenance of Chromosomes 1A (SMC1), Structural Maintenance Of Chromosomes 2 (SMC2), C. elegans tbp-1, D. melanogaster D1 protein, Salmonella typhimurium Hin recombinase, S. typhimurium Gin recombinase, S. typhimurium Pin recombinase, or S. typhimurium Cin recombinase. 265 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    42. The recombinant RT protein of any one of claims 38-41, wherein the linker: (a) a glycine-serine linker selected from the group consisting of (GS)n, (GSGGS)n, (SGGSG)n, (GGGS)n, (GGSG)n, (GGSGG)n, (GSGSG)n, (GSGGG)n, GGGSG)n, and (GSSSG)n, and wherein n represents an integer of at least 1; or (b) GGGS; or (c) SGGSG. 43. The recombinant RT protein of any one of claims 38-42, wherein the DNA binding domain is a S. cerevisiae datin (DAT1) DNA binding domain or fragment thereof. 44. The recombinant RT protein of any one of claims 38-43, wherein the DNA binding domain comprises: (a) an amino acid sequence selected from the group consisting of SEQ ID NO: 2, 3, 5, 6, 8, 9, and 11-24; or (b) a nucleic acid sequence of SEQ ID NO: 25. 45. The recombinant RT protein of any one of claims 38-44, wherein the RT polypeptide is selected from the group consisting of 42B L (SEQ ID NO: 145), 50A+G (SEQ ID NO: 147), SOLD 022 (SEQ ID NO: 105), SOLD 023 (SEQ ID NO: 107), SOLD 025 (SEQ ID NO: 111), SOLD 031 (SEQ ID NO: 123), SOLD 033 (SEQ ID NO: 127), SOLD 034 (SEQ ID NO: 129), SOLD 035 (SEQ ID NO: 131), SOLD 001 (SEQ ID NO: 65), SOLD 33 VDG (SEQ ID NO: 173), and an RT polypeptide set forth in SEQ ID NO: 143, SEQ ID NO: 172. 46. The recombinant RT protein of any one of claims 38-45 comprising, consisting essentially of, or consisting of SEQ ID NO: 174-188. 47. The engineered RT polypeptide of any one of claims 1-37, or the recombinant RT protein of any one of claims 38-46, wherein the engineered RT polypeptide or the recombinant RT protein further comprises a tag protein selected from the group consisting of an affinity tag, a fluorescent tag, or an expression and/or solubility enhancement tag. 266 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    48. The engineered RT polypeptide or the recombinant RT protein of claim 47, wherein the tag is selected from hexahistidine tag (his-tag), small ubiquitin-like modifier tag (SUMO), a VariFlex C-Terminal solubility enhancement tag, a short peptide C-terminal tag, Thioredoxin (Trx) tag, Solubility-enhancer peptide sequences (SET) tag, IgG domain B1 of Protein G (GB1) tag, IgG repeat domain ZZ of Protein A (ZZ) tag, Solubility enhancing Ubiquitous Tag (SNUT tag), Seventeen kilodalton protein (Skp tag), Phage T7 protein kinase (T7PK) tag, E. coli secreted protein A (EspA) tag, Monomeric bacteriophage T70.3 protein (Orc protein) (Mocr) tag, E. coli trypsin inhibitor (Ecotin) tag, Calcium-binding protein (CaBP) tag, Stress-responsive arsenate reductase (ArsC) tag, N-terminal fragment of translation initiation factor IF2 (IF2-domain I) tag, N- terminal fragment of translation initiation factor IF2 (Expressivity) tag, Fasciola hepatica 8-kDa antigen tag (Fh8), Glutathione-S-transferase (GST) tag, maltose-binding protein tag (MBP), Flag tag peptide (FLAG), streptavidin binding peptide tag (Strep-II; strep), calmodulin-binding protein tag (CBP), mutated dehalogenase tag (HaloTag), staphylococcal Protein A (Protein A), intein mediated purification with the chitin-binding domain (IMPACT (CBD)), cellulose-binding module (CBM), dockerin domain of Clostridium josui tag (Dock), fungal avidin-like protein (Tamavidin). 49. The engineered RT polypeptide or the recombinant RT protein of claim 47, wherein the tag is an affinity tag selected from hexahistidine tag (his-tag), Fasciola hepatica 8-kDa antigen tag (Fh8), Glutathione-S-transferase (GST) tag, maltose-binding protein tag (MBP), Flag tag peptide (FLAG), streptavidin binding peptide tag (Strep-II), calmodulin- binding protein tag (CBP), mutated dehalogenase tag (HaloTag), staphylococcal Protein A (Protein A), intein mediated purification with the chitin-binding domain (IMPACT (CBD)), cellulose-binding module (CBM), dockerin domain of Clostridium josui tag (Dock), fungal avidin-like protein (Tamavidin). 50. The engineered RT polypeptide or the recombinant RT protein of claim 47, wherein the engineered RT polypeptide or the recombinant RT protein comprises: (a) an hexahistidine tag (his-tag); or (b) an amino acid sequence of SEQ ID NO: 62; or an amino acid sequence having at least 90% sequence identity to the amino acid sequence of SEQ ID NO: 62. 267 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    51. The engineered RT polypeptide or the recombinant RT protein of claim 47, wherein the engineered RT polypeptide or the recombinant RT protein comprises a solubility enhancer tag selected from the group consisting of a SUMO tag, a GST tag, a Trx tag, a VariFlex C-Terminal solubility enhancement tag, a short peptide C-terminal tag, an Fh8 tag, MBP tag, SET tag, GB1 tag, ZZ tag, HaloTag, SNUT tag, Skp tag, T7PK tag, EspA tag, Mocr tag, Ecotin tag, CaBO tag, ArsC tag, IF2-domain I tag, Expressivity tag, RpoA, tag, SlyD, tag, Tsf tag, RpoS tag, PotD tag, Crr tag, msyB tag, yigD tag, and rpoD tag. 52. The engineered RT polypeptide or the recombinant RT protein of claim 47, wherein the engineered RT polypeptide or the recombinant RT protein comprises: (a) a short peptide C-terminal tag; (b) an amino acid sequence of SEQ ID NO: 193; or (c) an amino acid sequence having at least 90% sequence identity to the amino acid sequence of SEQ ID NO: 193. 53. The engineered RT polypeptide or the recombinant RT protein of claim 47, wherein the tag further comprises: (a) an endoprotein cleavage sequence; (b) a cleavage sequence recognized by an endoprotein selected from the group consisting of alanine carboxypeptidase, Armillaria mellea astacin, bacterial leucyl aminopeptidase, cancer procoagulant, cathepsin B, clostripain, cytosol alanyl aminopeptidase, elastase, endoproteinase Arg-C, enterokinase (EnTK), gastricsin, gelatinase, Gly-X carboxypeptidase, glycyl endopeptidase, human rhinovirus 3C protease, hypodermin C, Iga-specific serine endopeptidase, leucyl aminopeptidase, leucyl endopeptidase, lysC, lysosomal pro-X carboxypeptidase, lysyl aminopeptidase, methionyl aminopeptidase, myxobacter, nardilysin, pancreatic endopeptidase E, picornain 2A, picornain 3C, proendopeptidase, prolyl aminopeptidase, proprotein convertase I, proprotein convertase II, russellysin, saccharopepsin, semenogelase, T- plasminogen activator, thrombin (Thr), tissue kallikrein, tobacco etch virus (TEV), togavirin, tryptophanyl aminopeptidase, U-plasminogen activator, V8, venombin A, venombin AB, factor Xa (Xa), and Xaa-pro aminopeptidase; or 268 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    (c) an endoprotein cleavage sequence comprising the amino acid sequence of SEQ ID NO: 194, SEQ ID NO: 195, SEQ ID NO: 196, SEQ ID NO: 197, or SEQ ID NO: 198. 54. The engineered RT polypeptide of any one of claims 1-37 or 47-53, or the recombinant RT protein of any one of claims 38-52, wherein the engineered RT polypeptide or the recombinant RT protein exhibits increased template switching (TS) efficiency, increased processivity efficiency, increased binding affinity, increased transcription efficiency, increased chemical tolerance, improved ability to yield mitochondrial unique molecular identity (UMI) counts, improved ability to yield ribosomal unique molecular identity (UMI) counts, longer shelf life, higher strand displacement, higher end-to-end template jumping, or any combination thereof, when compared to an RT polypeptide or a recombinant RT protein lacking a conjugated DNA binding domain. 55. The engineered RT polypeptide of any one of claims 1-37 or 47-53, or the recombinant RT protein of any one of claims 38-53, wherein the engineered RT polypeptide or the recombinant RT protein comprises at least two or more of increased template switching (TS) efficiency, increased processivity efficiency, increased binding affinity, increased transcription efficiency, increased chemical tolerance, improved ability to yield mitochondrial unique molecular identity (UMI) counts, longer shelf life, higher strand displacement, higher end-to-end template jumping, or improved ability to yield ribosomal unique molecular identity (UMI) counts, when compared to an RT polypeptide or a recombinant RT protein lacking a conjugated DNA binding domain. 56. The engineered RT polypeptide of any one of claims 1-37 and 47-53 or the recombinant RT protein of any one of claims 38-53, wherein the recombinant RT protein or the engineered RT exhibits increased transcript capture during amplification, when compared to an RT polypeptide or a recombinant RT protein lacking a conjugated DNA binding domain. 57. The engineered RT polypeptide or the recombinant RT protein of claim 56, wherein the DNA binding domain enhances the hybridization of a transcript and a primer during a nucleic acid amplification process, when compared to an RT polypeptide or a recombinant RT protein lacking a conjugated DNA binding domain. 269 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    58. The engineered RT polypeptide or the recombinant RT protein of claim 56 or 57, wherein the primer comprises a poly-dT or a poly(dT)VN sequence and a non-poly(dT) sequence; and the transcript comprises a poly-dA sequence. 59. The engineered RT polypeptide or the recombinant RT protein of any one of claims 56- 58, wherein the DNA binding domain stabilizes the oligo(A)-oligo(T) based transcript- primer complex during a nucleic acid amplification process. 60. The engineered RT polypeptide or the recombinant RT protein of any one of claims 56- 59, wherein the primer is a barcoded molecule. 61. The engineered RT polypeptide or the recombinant RT protein of any one of claims 56- 60, wherein the transcript is a nucleic acid molecule selected from a RNA, a mRNA, or a DNA. 62. An isolated nucleic acid molecule encoding: (a) the engineered RT polypeptide of any one of claims 1-37, or 47-61; or (b) the recombinant RT protein of any one of claims 38-61. 63. The isolated nucleic acid molecule of claim 62, wherein the nucleic acid molecule comprises a sequence selected from SEQ ID NO: 25, SEQ ID NO: 136, SEQ ID NO:138, SEQ ID NO:140, SEQ ID NO:142, SEQ ID NO:144, SEQ ID NO:146, SEQ ID NO:148, SEQ ID NO:150, SEQ ID NO:152, SEQ ID NO:154, SEQ ID NO:156, SEQ ID NO:158, SEQ ID NO:160, SEQ ID NO:162, SEQ ID NO:164, SEQ ID NO:167, SEQ ID NO:169, or SEQ ID NO: 171; or a nucleic acid sequence of Table 2. 64. An expression vector comprising the isolated nucleic acid of claim 62 or 63. 65. A host cell transfected with the expression vector of claim 64 or the isolated nucleic acid of claim 62 or 63. 66. A composition comprising: (a) the recombinant RT protein of any one of claims 38-61; or (b) the engineered RT polypeptide of any one of claims 1-37 or 47-61; or (d) an expression vector of claim 64; or 270 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    (e) a host cell of claim 65; and (f) a buffer. 67. A method for performing a reverse transcription reaction for generating a nucleic acid product from an RNA template comprising contacting under suitable conditions a biological sample or extract thereof with an engineered RT polypeptide of any of claims 1-37 or 47-61, or the recombinant RT protein of any one of claims 38-61. 68. The method of claim 67, wherein the biological sample comprises a cell, optionally wherein the cell is permeabilized and/or optionally wherein the cell is fixed. 69. The method of claim 67, wherein the biological sample comprises a cell bead, optionally wherein the cell bead is fixed. 70. The method of claim 67, wherein the biological sample comprises a nucleus, optionally wherein the nucleus is permeabilized and optionally wherein the nucleus is fixed. 71. The method of claim 67, wherein the biological sample comprises (a) a suitable cellular preparation selected from cell populations and/or single cells, or (b) a tissue. 72. The method of claim 71, wherein the sample comprises cells in suspension, fresh cells, fixed cells, or cells and tissues immobilized on various solid surfaces. 73. The method of claim 67, wherein the biological sample is a cell, a cell bead, or a nucleus, and the reverse transcription reaction is part of a single cell RNA sequencing assay. 74. The method of claim 73, wherein the single cell RNA sequencing assay further comprises, prior to the reverse transcription, partitioning the cell, cell bead, or nucleus into a partition. 75. The method of claim 74, wherein the single cell RNA sequencing assay further comprises, after the reverse transcription reaction, hybridizing the nucleic acid product to an oligonucleotide molecule comprising a partition-specific barcode. 76. The method of claim 67, wherein the biological sample is a cell or tissue sample immobilized on a surface, and the reverse transcription reaction is part of a spatial RNA sequencing assay. 271 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    77. The method of any one of claims 67-76, wherein the engineered RT polypeptide or the recombinant RT protein enhances template switching (TS) efficiency, processivity efficiency, binding affinity, transcription efficiency, chemical tolerance, ability to yield mitochondrial unique molecular identity (UMI) counts, ability to yield ribosomal unique molecular identity (UMI) counts, strand displacement, end-to-end template jumping, or any combination thereof, when compared to an RT polypeptide or a recombinant RT protein lacking a conjugated DNA binding domain. 78. The method of any one of claims 67-76, wherein the engineered RT polypeptide or the recombinant RT protein enhances at least two or more of template switching (TS) efficiency, processivity efficiency, binding affinity, transcription efficiency, chemical tolerance, ability to yield mitochondrial unique molecular identity (UMI) counts, strand displacement, end-to-end template jumping, or ability to yield ribosomal unique molecular identity (UMI) counts, when compared to an RT polypeptide or a recombinant RT protein lacking a conjugated DNA binding domain. 79. The method of any one of claims 67-76, wherein the recombinant RT protein or the engineered RT enhances transcript capture during amplification, when compared to an RT polypeptide or a recombinant RT protein lacking a conjugated DNA binding domain. 80. The method of any one of claims 67-79, wherein the DNA binding domain of the recombinant RT protein or the engineered RT enhances the hybridization of a transcript and a primer during the amplification process, when compared to an RT polypeptide or a recombinant RT protein lacking a conjugated DNA binding domain. 81. The method of any one of claims 67-80, wherein: the engineered RT polypeptide or the recombinant RT protein comprises: (a) a DNA binding domain comprising an amino acid sequence selected from SEQ ID NO:2, 3, 5, 6, 8, 9, or 11-24; and (b) an amino acid sequence selected from SEQ ID NOs: 27-61, 63, 65, 67, 69, 71, 73, 75, 77, 79, 81, 83, 85, 87, 89, 91, 93, 95, 97, 99, 101, 103, 105, 107, 109, 111, 113, 115, 117, 119, 121, 123, 125, 127, 129, 131, 133, 141, 143, 145, 147, 149, 151, 157, 159, 172, or 173. 272 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    82. The method of any one of claims67-81, wherein the amino acid sequence of the engineered RT polypeptide or the recombinant RT protein comprises an amino acid sequence having at least about 90%, at least about 92%, at least about 95%, at least about 97%, at least about 98%, or at least about 99% sequence identity to SEQ ID NO: 174- 188. 83. The method of any one of claims67-82, wherein the engineered RT polypeptide or the recombinant RT protein comprises: (a) M39V, M66I, Q91R, I347V, and H594Q substitution in SEQ ID NO: 143; (b) SEQ ID NO: 129 (SOLD 034); (c) SOLD 001 (SEQ ID NO: 65); or (d) SOLD 33 VDG (SEQ ID NO: 173). 84. The method of any one of claims 67-83, wherein the engineered RT polypeptide or the recombinant RT protein comprises M39V, T542D, D583N, E607G, A644V, D653H, K658R, and L671P in SEQ ID NO: 1 or 143. 85. The method of any one of claims 67-84, wherein the engineered RT polypeptide or the recombinant RT protein comprises: (a) M39V, T542D, D583N, E607G, A644V, D653H, K658R, L671P in SEQ ID NO: 143; or (b) SEQ ID NO: 111 (SOLD 025). 86. The method of any one of claims 67-85, wherein the engineered RT or the recombinant RT protein comprises a M39V, M66I, Q91R, I347V, H594Q in SEQ ID NO: 143. 87. The method of any one of claims 67-86, wherein the engineered RT polypeptide or the recombinant RT protein comprises an amino acid sequence that is at least about 90%, at least about 92%, at least about 95%, at least about 97%, at least about 98%, or at least about 99% identical to an amino acid sequence disclosed in Table 1 or Table 2. 88. A method of using the engineered RT polypeptide of any one of claims 1-37 or 47-61, or the recombinant RT protein of any one of claims 38-61, the method comprising contacting the engineered RT polypeptide or the recombinant RT protein with a nucleic 273 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    acid template under suitable conditions to produce a polymerized nucleic acid product, wherein the nucleic acid template comprises an RNA 89. A nucleic acid extension method comprising: (a) contacting a target nucleic acid molecule with an engineered reverse transcriptase polypeptide or a recombinant RT protein and a plurality of nucleic acid barcoded molecules comprising a barcode sequence, and (b) incubating the target nucleic acid, the engineered RT polypeptide or the recombinant RT protein and barcoded molecules under suitable conditions in which the barcoded molecules are extended by the engineered RT polypeptide or the recombinant RT protein, wherein the engineered RT polypeptide comprises the amino acid sequence of an engineered RT polypeptide of any one of claims 1-37 or 47-61, or a recombinant RT protein of any one of claims 38-61. 90. The method of any one of claims 67-89, wherein the recombinant RT protein or the engineered RT polypeptide exhibits increased transcript capture during amplification. 91. The method of any one of claims67-90, wherein the DNA binding domain enhances the hybridization of a transcript and a primer during a nucleic acid amplification process. 92. The method of any one of claims 67-91, wherein the primer comprises a poly-dT or a poly(dT)VN sequence and a non-poly(dT) sequence; and the transcript comprises a poly- dA sequence. 93. The method of any one of claims 67-92, wherein the DNA binding domain stabilizes the oligo(A)-olgo(T) based transcript-primer complex during a nucleic acid amplification process. 94. The method of any one of claims 67-93, wherein the primer is a barcoded molecule. 95. The method of any one of claims 67-94, wherein the recombinant RT protein or the engineered RT polypeptide performs the first strand complementary DNA (cDNA) reaction. 96. The method of claim 95, wherein the first strand cDNA is amplified using a DNA polymerase to generate a second strand cDNA. 274 4876-6828-0003.1 Foley Ref.: 131488-0215 10X Genomics Ref.: 100-165501PC    97. A kit comprising: (a) a recombinant RT protein of any one of claims 38-61; or (b) an engineered reverse transcriptase polypeptide of any one of claims 1-37 or 47- 61; or (c) the isolated nucleic acid of claim 62 or 63; or (d) an expression vector of claim 64; or (e) a host cell of claim 65; or (f) the composition of claim 66; and (g) instructions. 275 4876-6828-0003.1
PCT/US2024/033519 2023-06-13 2024-06-12 Sequence specific dna binding proteins Ceased WO2024258911A1 (en)

Priority Applications (1)

Application Number Priority Date Filing Date Title
EP24739890.2A EP4728062A1 (en) 2023-06-13 2024-06-12 Sequence specific dna binding proteins

Applications Claiming Priority (4)

Application Number Priority Date Filing Date Title
US202363472726P 2023-06-13 2023-06-13
US63/472,726 2023-06-13
US202463622402P 2024-01-18 2024-01-18
US63/622,402 2024-01-18

Publications (1)

Publication Number Publication Date
WO2024258911A1 true WO2024258911A1 (en) 2024-12-19

Family

ID=91853433

Family Applications (1)

Application Number Title Priority Date Filing Date
PCT/US2024/033519 Ceased WO2024258911A1 (en) 2023-06-13 2024-06-12 Sequence specific dna binding proteins

Country Status (2)

Country Link
EP (1) EP4728062A1 (en)
WO (1) WO2024258911A1 (en)

Citations (56)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
US20030198944A1 (en) 1997-04-22 2003-10-23 Invitrogen Corporation Compositions and methods for reverse transcription of nucleic acid molecules
US7709198B2 (en) 2005-06-20 2010-05-04 Advanced Cell Diagnostics, Inc. Multiplex detection of nucleic acids
US20130171621A1 (en) 2010-01-29 2013-07-04 Advanced Cell Diagnostics Inc. Methods of in situ detection of nucleic acids
US20140378345A1 (en) 2012-08-14 2014-12-25 10X Technologies, Inc. Compositions and methods for sample processing
US20150376609A1 (en) 2014-06-26 2015-12-31 10X Genomics, Inc. Methods of Analyzing Nucleic Acids from Individual Cells or Cell Populations
WO2016057552A1 (en) 2014-10-06 2016-04-14 The Board Of Trustees Of The Leland Stanford Junior University Multiplexed detection and quantification of nucleic acids in single-cells
US9593365B2 (en) 2012-10-17 2017-03-14 Spatial Transcriptions Ab Methods and product for optimising localised or spatial detection of gene expression in a tissue sample
US9727810B2 (en) 2015-02-27 2017-08-08 Cellular Research, Inc. Spatially addressable molecular barcoding
WO2017144338A1 (en) 2016-02-22 2017-08-31 Miltenyi Biotec Gmbh Automated analysis tool for biological specimens
US9783841B2 (en) 2012-10-04 2017-10-10 The Board Of Trustees Of The Leland Stanford Junior University Detection of target nucleic acids in a cellular sample
US9879313B2 (en) 2013-06-25 2018-01-30 Prognosys Biosciences, Inc. Methods and systems for determining spatial patterns of biological targets in a sample
US20180105808A1 (en) 2016-10-19 2018-04-19 10X Genomics, Inc. Methods and systems for barcoding nucleic acid molecules from individual cells or cell populations
WO2018091676A1 (en) 2016-11-17 2018-05-24 Spatial Transcriptomics Ab Method for spatial tagging and analysing nucleic acids in a biological specimen
US10030261B2 (en) 2011-04-13 2018-07-24 Spatial Transcriptomics Ab Method and product for localized or spatial detection of nucleic acid in a tissue sample
US10041949B2 (en) 2013-09-13 2018-08-07 The Board Of Trustees Of The Leland Stanford Junior University Multiplexed imaging of tissues using mass tags and secondary ion mass spectrometry
US10059990B2 (en) 2015-04-14 2018-08-28 Massachusetts Institute Of Technology In situ nucleic acid sequencing of expanded biological samples
US20190085383A1 (en) 2014-07-11 2019-03-21 President And Fellows Of Harvard College Methods for High-Throughput Labelling and Detection of Biological Features In Situ Using Microscopy
US10317321B2 (en) 2015-08-07 2019-06-11 Massachusetts Institute Of Technology Protein retention expansion microscopy
US10323278B2 (en) 2016-12-22 2019-06-18 10X Genomics, Inc. Methods and systems for processing polynucleotides
US10364457B2 (en) 2015-08-07 2019-07-30 Massachusetts Institute Of Technology Nanoscale imaging of proteins and nucleic acids via expansion microscopy
US10400280B2 (en) 2012-08-14 2019-09-03 10X Genomics, Inc. Methods and systems for processing polynucleotides
US20190330617A1 (en) 2016-08-31 2019-10-31 President And Fellows Of Harvard College Methods of Generating Libraries of Nucleic Acid Sequences for Detection via Fluorescent in Situ Sequ
US10480022B2 (en) 2010-04-05 2019-11-19 Prognosys Biosciences, Inc. Spatially encoded biological assays
US10494662B2 (en) 2013-03-12 2019-12-03 President And Fellows Of Harvard College Method for generating a three-dimensional nucleic acid containing matrix
US20200032335A1 (en) 2018-07-27 2020-01-30 10X Genomics, Inc. Systems and methods for metabolome analysis
WO2020047010A2 (en) 2018-08-28 2020-03-05 10X Genomics, Inc. Increasing spatial array resolution
WO2020047005A1 (en) 2018-08-28 2020-03-05 10X Genomics, Inc. Resolving spatial arrays
US20200080136A1 (en) 2016-09-22 2020-03-12 William Marsh Rice University Molecular hybridization probes for complex sequence capture and analysis
WO2020053655A1 (en) 2018-09-13 2020-03-19 Zenith Epigenetics Ltd. Combination therapy for the treatment of triple-negative breast cancer
US10640816B2 (en) 2015-07-17 2020-05-05 Nanostring Technologies, Inc. Simultaneous quantification of gene expression in a user-defined region of a cross-sectioned tissue
WO2020123320A2 (en) 2018-12-10 2020-06-18 10X Genomics, Inc. Imaging system hardware
US20200224244A1 (en) 2017-10-06 2020-07-16 Cartana Ab Rna templated ligation
US10724078B2 (en) 2015-04-14 2020-07-28 Koninklijke Philips N.V. Spatial mapping of molecular profiles of biological tissue samples
US20200239946A1 (en) 2017-10-11 2020-07-30 Expansion Technologies Multiplexed in situ hybridization of tissue sections for spatially resolved transcriptomics with expansion microscopy
US20200256867A1 (en) 2016-12-09 2020-08-13 Ultivue, Inc. Methods for Multiplex Imaging Using Labeled Nucleic Acid Imaging Agents
WO2020176788A1 (en) 2019-02-28 2020-09-03 10X Genomics, Inc. Profiling of biological analytes with spatially barcoded oligonucleotide arrays
US10774374B2 (en) 2015-04-10 2020-09-15 Spatial Transcriptomics AB and Illumina, Inc. Spatially distinguished, multiplex nucleic acid analysis of biological specimens
US10913975B2 (en) 2015-07-27 2021-02-09 Illumina, Inc. Spatial mapping of nucleic acid sequence information
US10995361B2 (en) 2017-01-23 2021-05-04 Massachusetts Institute Of Technology Multiplexed signal amplified FISH via splinted ligation amplification and sequencing
US20210140982A1 (en) 2019-10-18 2021-05-13 10X Genomics, Inc. Identification of spatial biomarkers of brain disorders and methods of using the same
US11008608B2 (en) 2016-02-26 2021-05-18 The Board Of Trustees Of The Leland Stanford Junior University Multiplexed single molecule RNA visualization with a two-probe proximity ligation system
US20210150707A1 (en) 2019-11-18 2021-05-20 10X Genomics, Inc. Systems and methods for binary tissue classification
US20210155982A1 (en) 2019-11-21 2021-05-27 10X Genomics, Inc. Pipeline for spatial analysis of analytes
US20210158522A1 (en) 2019-11-22 2021-05-27 10X Genomics, Inc. Systems and methods for spatial analysis of analytes using fiducial alignment
US20210189475A1 (en) 2018-12-10 2021-06-24 10X Genomics, Inc. Imaging system hardware
US20210198741A1 (en) 2019-12-30 2021-07-01 10X Genomics, Inc. Identification of spatial biomarkers of heart disorders and methods of using the same
WO2021133849A1 (en) 2019-12-23 2021-07-01 10X Genomics, Inc. Methods for spatial analysis using rna-templated ligation
US20210199660A1 (en) 2019-11-22 2021-07-01 10X Genomics, Inc. Biomarkers of breast cancer
US11104936B2 (en) 2014-04-18 2021-08-31 William Marsh Rice University Competitive compositions of nucleic acid molecules for enrichment of rare-allele-bearing species
US11168350B2 (en) 2016-07-27 2021-11-09 The Board Of Trustees Of The Leland Stanford Junior University Highly-multiplexed fluorescent imaging
WO2021252747A1 (en) 2020-06-10 2021-12-16 1Ox Genomics, Inc. Fluid delivery methods
WO2022061152A2 (en) 2020-09-18 2022-03-24 10X Genomics, Inc. Sample handling apparatus and fluid delivery methods
US11352667B2 (en) 2016-06-21 2022-06-07 10X Genomics, Inc. Nucleic acid sequencing
WO2022140028A1 (en) 2020-12-21 2022-06-30 10X Genomics, Inc. Methods, compositions, and systems for capturing probes and/or barcodes
US11447807B2 (en) 2016-08-31 2022-09-20 President And Fellows Of Harvard College Methods of combining the detection of biomolecules into a single assay using fluorescent in situ sequencing
WO2022232571A1 (en) * 2021-04-30 2022-11-03 10X Genomics, Inc. Fusion rt variants for improved performance

Patent Citations (68)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
US20030198944A1 (en) 1997-04-22 2003-10-23 Invitrogen Corporation Compositions and methods for reverse transcription of nucleic acid molecules
US7709198B2 (en) 2005-06-20 2010-05-04 Advanced Cell Diagnostics, Inc. Multiplex detection of nucleic acids
US8604182B2 (en) 2005-06-20 2013-12-10 Advanced Cell Diagnostics, Inc. Multiplex detection of nucleic acids
US8951726B2 (en) 2005-06-20 2015-02-10 Advanced Cell Diagnostics, Inc. Multiplex detection of nucleic acids
US20130171621A1 (en) 2010-01-29 2013-07-04 Advanced Cell Diagnostics Inc. Methods of in situ detection of nucleic acids
US10480022B2 (en) 2010-04-05 2019-11-19 Prognosys Biosciences, Inc. Spatially encoded biological assays
US10030261B2 (en) 2011-04-13 2018-07-24 Spatial Transcriptomics Ab Method and product for localized or spatial detection of nucleic acid in a tissue sample
US20140378345A1 (en) 2012-08-14 2014-12-25 10X Technologies, Inc. Compositions and methods for sample processing
US10400280B2 (en) 2012-08-14 2019-09-03 10X Genomics, Inc. Methods and systems for processing polynucleotides
US9783841B2 (en) 2012-10-04 2017-10-10 The Board Of Trustees Of The Leland Stanford Junior University Detection of target nucleic acids in a cellular sample
US9593365B2 (en) 2012-10-17 2017-03-14 Spatial Transcriptions Ab Methods and product for optimising localised or spatial detection of gene expression in a tissue sample
US10494662B2 (en) 2013-03-12 2019-12-03 President And Fellows Of Harvard College Method for generating a three-dimensional nucleic acid containing matrix
US9879313B2 (en) 2013-06-25 2018-01-30 Prognosys Biosciences, Inc. Methods and systems for determining spatial patterns of biological targets in a sample
US10041949B2 (en) 2013-09-13 2018-08-07 The Board Of Trustees Of The Leland Stanford Junior University Multiplexed imaging of tissues using mass tags and secondary ion mass spectrometry
US11104936B2 (en) 2014-04-18 2021-08-31 William Marsh Rice University Competitive compositions of nucleic acid molecules for enrichment of rare-allele-bearing species
US20150376609A1 (en) 2014-06-26 2015-12-31 10X Genomics, Inc. Methods of Analyzing Nucleic Acids from Individual Cells or Cell Populations
US20190085383A1 (en) 2014-07-11 2019-03-21 President And Fellows Of Harvard College Methods for High-Throughput Labelling and Detection of Biological Features In Situ Using Microscopy
WO2016057552A1 (en) 2014-10-06 2016-04-14 The Board Of Trustees Of The Leland Stanford Junior University Multiplexed detection and quantification of nucleic acids in single-cells
US9727810B2 (en) 2015-02-27 2017-08-08 Cellular Research, Inc. Spatially addressable molecular barcoding
US10002316B2 (en) 2015-02-27 2018-06-19 Cellular Research, Inc. Spatially addressable molecular barcoding
US10774374B2 (en) 2015-04-10 2020-09-15 Spatial Transcriptomics AB and Illumina, Inc. Spatially distinguished, multiplex nucleic acid analysis of biological specimens
US10724078B2 (en) 2015-04-14 2020-07-28 Koninklijke Philips N.V. Spatial mapping of molecular profiles of biological tissue samples
US10059990B2 (en) 2015-04-14 2018-08-28 Massachusetts Institute Of Technology In situ nucleic acid sequencing of expanded biological samples
US10640816B2 (en) 2015-07-17 2020-05-05 Nanostring Technologies, Inc. Simultaneous quantification of gene expression in a user-defined region of a cross-sectioned tissue
US10913975B2 (en) 2015-07-27 2021-02-09 Illumina, Inc. Spatial mapping of nucleic acid sequence information
US10364457B2 (en) 2015-08-07 2019-07-30 Massachusetts Institute Of Technology Nanoscale imaging of proteins and nucleic acids via expansion microscopy
US10317321B2 (en) 2015-08-07 2019-06-11 Massachusetts Institute Of Technology Protein retention expansion microscopy
WO2017144338A1 (en) 2016-02-22 2017-08-31 Miltenyi Biotec Gmbh Automated analysis tool for biological specimens
US11008608B2 (en) 2016-02-26 2021-05-18 The Board Of Trustees Of The Leland Stanford Junior University Multiplexed single molecule RNA visualization with a two-probe proximity ligation system
US11352667B2 (en) 2016-06-21 2022-06-07 10X Genomics, Inc. Nucleic acid sequencing
US11168350B2 (en) 2016-07-27 2021-11-09 The Board Of Trustees Of The Leland Stanford Junior University Highly-multiplexed fluorescent imaging
US11447807B2 (en) 2016-08-31 2022-09-20 President And Fellows Of Harvard College Methods of combining the detection of biomolecules into a single assay using fluorescent in situ sequencing
US20190330617A1 (en) 2016-08-31 2019-10-31 President And Fellows Of Harvard College Methods of Generating Libraries of Nucleic Acid Sequences for Detection via Fluorescent in Situ Sequ
US20200080136A1 (en) 2016-09-22 2020-03-12 William Marsh Rice University Molecular hybridization probes for complex sequence capture and analysis
US20180105808A1 (en) 2016-10-19 2018-04-19 10X Genomics, Inc. Methods and systems for barcoding nucleic acid molecules from individual cells or cell populations
WO2018091676A1 (en) 2016-11-17 2018-05-24 Spatial Transcriptomics Ab Method for spatial tagging and analysing nucleic acids in a biological specimen
US20200256867A1 (en) 2016-12-09 2020-08-13 Ultivue, Inc. Methods for Multiplex Imaging Using Labeled Nucleic Acid Imaging Agents
US10323278B2 (en) 2016-12-22 2019-06-18 10X Genomics, Inc. Methods and systems for processing polynucleotides
US10995361B2 (en) 2017-01-23 2021-05-04 Massachusetts Institute Of Technology Multiplexed signal amplified FISH via splinted ligation amplification and sequencing
US20200224244A1 (en) 2017-10-06 2020-07-16 Cartana Ab Rna templated ligation
US20200239946A1 (en) 2017-10-11 2020-07-30 Expansion Technologies Multiplexed in situ hybridization of tissue sections for spatially resolved transcriptomics with expansion microscopy
US20200032335A1 (en) 2018-07-27 2020-01-30 10X Genomics, Inc. Systems and methods for metabolome analysis
WO2020047005A1 (en) 2018-08-28 2020-03-05 10X Genomics, Inc. Resolving spatial arrays
WO2020047004A2 (en) 2018-08-28 2020-03-05 10X Genomics, Inc. Methods of generating an array
WO2020047007A2 (en) 2018-08-28 2020-03-05 10X Genomics, Inc. Methods for generating spatially barcoded arrays
WO2020047010A2 (en) 2018-08-28 2020-03-05 10X Genomics, Inc. Increasing spatial array resolution
WO2020047002A1 (en) 2018-08-28 2020-03-05 10X Genomics, Inc. Method for transposase-mediated spatial tagging and analyzing genomic dna in a biological sample
WO2020053655A1 (en) 2018-09-13 2020-03-19 Zenith Epigenetics Ltd. Combination therapy for the treatment of triple-negative breast cancer
US20200277663A1 (en) 2018-12-10 2020-09-03 10X Genomics, Inc. Methods for determining a location of a biological analyte in a biological sample
US20210189475A1 (en) 2018-12-10 2021-06-24 10X Genomics, Inc. Imaging system hardware
WO2020123320A2 (en) 2018-12-10 2020-06-18 10X Genomics, Inc. Imaging system hardware
WO2020176788A1 (en) 2019-02-28 2020-09-03 10X Genomics, Inc. Profiling of biological analytes with spatially barcoded oligonucleotide arrays
US20210140982A1 (en) 2019-10-18 2021-05-13 10X Genomics, Inc. Identification of spatial biomarkers of brain disorders and methods of using the same
WO2021102003A1 (en) 2019-11-18 2021-05-27 10X Genomics, Inc. Systems and methods for tissue classification
US20210150707A1 (en) 2019-11-18 2021-05-20 10X Genomics, Inc. Systems and methods for binary tissue classification
WO2021102039A1 (en) 2019-11-21 2021-05-27 10X Genomics, Inc, Spatial analysis of analytes
US20210155982A1 (en) 2019-11-21 2021-05-27 10X Genomics, Inc. Pipeline for spatial analysis of analytes
US20210199660A1 (en) 2019-11-22 2021-07-01 10X Genomics, Inc. Biomarkers of breast cancer
US20210158522A1 (en) 2019-11-22 2021-05-27 10X Genomics, Inc. Systems and methods for spatial analysis of analytes using fiducial alignment
WO2021102005A1 (en) 2019-11-22 2021-05-27 10X Genomics, Inc. Systems and methods for spatial analysis of analytes using fiducial alignment
WO2021133849A1 (en) 2019-12-23 2021-07-01 10X Genomics, Inc. Methods for spatial analysis using rna-templated ligation
US11505828B2 (en) 2019-12-23 2022-11-22 10X Genomics, Inc. Methods for spatial analysis using RNA-templated ligation
US11332790B2 (en) 2019-12-23 2022-05-17 10X Genomics, Inc. Methods for spatial analysis using RNA-templated ligation
US20210198741A1 (en) 2019-12-30 2021-07-01 10X Genomics, Inc. Identification of spatial biomarkers of heart disorders and methods of using the same
WO2021252747A1 (en) 2020-06-10 2021-12-16 1Ox Genomics, Inc. Fluid delivery methods
WO2022061152A2 (en) 2020-09-18 2022-03-24 10X Genomics, Inc. Sample handling apparatus and fluid delivery methods
WO2022140028A1 (en) 2020-12-21 2022-06-30 10X Genomics, Inc. Methods, compositions, and systems for capturing probes and/or barcodes
WO2022232571A1 (en) * 2021-04-30 2022-11-03 10X Genomics, Inc. Fusion rt variants for improved performance

Non-Patent Citations (35)

* Cited by examiner, † Cited by third party
Title
ALTSCHUL ET AL., NAT'L CENT. BIOTECHNOL. INF.
ALTSCHUL ET AL., NUCLEIC ACIDS RES., vol. 25, 1997, pages 3389 - 3402
ATHEY ET AL., BMC BIOIN FORMATICS, vol. 18, 2017, pages 391 - 401
AUSUBEL ET AL.: "Current Protocols in Molecular Biology", vol. 22, 1994, JOHN WILEY & SONS, INC, article "Gene Expression in Recombinant Microorganisms"
BEBENEK: "A minor groove binding track in reverse transcriptase", 1 January 1997 (1997-01-01), XP093210585, Retrieved from the Internet <URL:https://www.nature.com/articles/nsb0397-194.pdf> *
BOSWORTH ET AL., NATURE, vol. 341, 1989, pages 167 - 168
BROSIUS ET AL., VIRUS GENES, vol. 11, 1995, pages 163 - 79
CHEN ET AL., SCIENCE, vol. 348, no. 6233, 2015
CHENG CHIEN-CHUNG ET AL: "Design and Characterization of a Short HMG-I/ DAT1 Peptide that Binds Specifically to the Minor Groove of DNA", JOURNAL OF THE CHINESE CHEMICAL SOCIETY., vol. 45, no. 5, 1 October 1998 (1998-10-01), CHINA, pages 619 - 624, XP093210721, ISSN: 0009-4536, DOI: 10.1002/jccs.199800093 *
COSTA ET AL., FRONT. MICROBIOL., vol. 63, no. 5, 2014
CREDLE ET AL., NUCLEIC ACIDS RES., vol. 45, no. 14, 21 August 2017 (2017-08-21), pages e128
CURRENT PROTOCOLS IN MOLECULAR BIOLOGY, 1987
ERGIN B ET AL., J PROTEOME RES., vol. 9, no. 10, 1 October 2010 (2010-10-01), pages 5188 - 96
ESPOSITOCHATTERJEE, CURR. OPIN. BIOTECHNOL., vol. 17, 2006, pages 353 - 358
FISHER TIMOTHY S. ET AL: "Mutations Proximal to the Minor Groove-Binding Track of Human Immunodeficiency Virus Type 1 Reverse Transcriptase Differentially Affect Utilization of RNA versus DNA as Template", JOURNAL OF VIROLOGY, vol. 77, no. 10, 15 May 2003 (2003-05-15), US, pages 5837 - 5845, XP093210621, ISSN: 0022-538X, Retrieved from the Internet <URL:https://www.ncbi.nlm.nih.gov/pmc/articles/PMC154037/pdf/2438.pdf> DOI: 10.1128/JVI.77.10.5837-5845.2003 *
GAO ET AL., BMC BIOL, vol. 15, 2017, pages 50
GUPTA ET AL., NATURE BIOTECHNOL., vol. 36, 2018, pages 1197 - 1202
JAMUR ET AL., METHOD MOL. BIOL., vol. 588, pages 63 - 66
KAP M. ET AL., PLOS ONE.;, vol. 6, no. 11, 2011, pages e27704
LATHAM GARY J. ET AL: "Vertical-scanning Mutagenesis of a Critical Tryptophan in the "Minor Groove Binding Track" of HIV-1 Reverse Transcriptase", JOURNAL OF BIOLOGICAL CHEMISTRY, vol. 275, no. 20, 1 May 2000 (2000-05-01), US, pages 15025 - 15033, XP093210664, ISSN: 0021-9258, DOI: 10.1074/jbc.M000279200 *
LEE ET AL., NAT. PROTOC., vol. 10, no. 3, 2015, pages 442 - 458
LEVIN, CELL, vol. 88, 1997, pages 5 - 8
MALHOTRA, A: "Guide to Protein Purification", vol. 463, 2009, ELSEVIER, article "Tagging for protein expression", pages: 239 - 258
MATHIESON W. ET AL., AM J CLIN PATHOL, vol. 146, no. 1, 2016, pages 25 - 40
NAJMUDIN S ET AL: "Crystal structures of an N-terminal fragment from moloney murine leukemia virus reverse transcriptase complexed with nucleic acid: functional implications for template-primer binding to the fingers domain", JOURNAL OF MOLECULAR BIOLOGY, ACADEMIC PRESS, UNITED KINGDOM, vol. 296, no. 2, 19 February 2000 (2000-02-19), pages 613 - 632, XP004461564, ISSN: 0022-2836, DOI: 10.1006/JMBI.1999.3477 *
OSCORBIN ET AL., FEBS LETT., vol. 594, 2020, pages 4338
PEARSON ET AL., PROC. NATL ACAD, vol. 85, 1988, pages 2444 - 2448
REARDON B.J ET AL: "<mark>DNA binding properties of the Saccharomyces cerevisiae DAT1 gene product</mark>", NUCLEIC ACIDS RESEARCH, vol. 23, no. 23, 1 January 1995 (1995-01-01), GB, pages 4900 - 4906, XP093210322, ISSN: 0305-1048, Retrieved from the Internet <URL:https://www.ncbi.nlm.nih.gov/pmc/articles/PMC307481/pdf/nar00023-0166.pdf> DOI: 10.1093/nar/23.23.4900 *
REARDON ET AL., NUCLEIC ACIDS RESEARCH, vol. 23, 1995, pages 4900
REARDON ET AL., PNAS, vol. 90, 1993, pages 11327
RODRIQUES ET AL., SCIENCE, vol. 363, no. 6434, 2019, pages 1463 - 1467
SEALEY DAVID C. ET AL: "The N-terminus of hTERT contains a DNA-binding domain and is required for telomerase activity and cellular immortalization", NUCLEIC ACIDS RESEARCH, vol. 38, no. 6, 24 December 2009 (2009-12-24), GB, pages 2019 - 2035, XP093208256, ISSN: 0305-1048, Retrieved from the Internet <URL:https://academic.oup.com/nar/article-pdf/38/6/2019/16770445/gkp1160.pdf> DOI: 10.1093/nar/gkp1160 *
THOMAS H. PETERSEN ET AL: "Utility of Telomerase-pot1 Fusion Protein in Vascular Tissue Engineering", CELL TRANSPLANTATION, vol. 19, no. 1, 1 January 2010 (2010-01-01), US, pages 79 - 87, XP055671385, ISSN: 0963-6897, DOI: 10.3727/096368909X478650 *
TREJO ET AL., PLOS ONE, vol. 14, no. 2, 2019
WINTERVARSHAVSKY, EMBO J, vol. 8, 1989, pages 1867

Also Published As

Publication number Publication date
EP4728062A1 (en) 2026-04-22

Similar Documents

Publication Publication Date Title
US20250257393A1 (en) Methods and compositions for detecting nucleic acids in a fixed biological sample
US20240254475A1 (en) Proteomic analysis with nucleic acid identifiers
JP6457564B2 (en) Proximity extension assay using exonuclease
EP4490319A1 (en) Molecular barcode readers for analyte detection
US8871686B2 (en) Methods of identifying a pair of binding partners
US9902993B2 (en) Hyperthermophilic polymerase enabled proximity extension assay
CN103154266B (en) Block reagent and using method thereof
CN115698318B (en) Control for proximity detection assay
US20160060687A1 (en) Methods and compositions to identify, quantify, and characterize target analytes and binding moieties
WO2025072119A1 (en) Methods, compositions, and kits for detecting spatial chromosomal interactions
CN111549177A (en) gRNA and kit for detecting SARS-CoV-2
CN112534049A (en) Method for processing nucleic acid samples
US20250277259A1 (en) Method for mapping rolling circle amplification products
US20240368567A1 (en) Recombinant reverse transcriptase variants for improved performance
US20250340933A1 (en) Methods, compositions, and kits for determining the presence and/or location of an exogenous target nucleic acid in a biological sample
Maranhao et al. An improved and readily available version of Bst DNA Polymerase for LAMP, and applications to COVID-19 diagnostics
WO2022265965A1 (en) Reverse transcriptase variants for improved performance
WO2021199770A1 (en) Method for detecting target nucleic acid
EP4728062A1 (en) Sequence specific dna binding proteins
WO2024238992A1 (en) Engineered non-strand displacing family b polymerases for reverse transcription and gap-fill applications
US20230374475A1 (en) Engineered thermophilic reverse transcriptase
KR102684067B1 (en) Cas complex for simultaneous detection of DNA and RNA and uses thereof
US20240271187A1 (en) Hybridisation-based sensor systems and probes
CN117693582A (en) Reverse transcriptase variants for improved performance
US20240228989A1 (en) Reverse transcriptase variants for improved performance

Legal Events

Date Code Title Description
121 Ep: the epo has been informed by wipo that ep was designated in this application

Ref document number: 24739890

Country of ref document: EP

Kind code of ref document: A1

WWE Wipo information: entry into national phase

Ref document number: 2024739890

Country of ref document: EP

NENP Non-entry into the national phase

Ref country code: DE

ENP Entry into the national phase

Ref document number: 2024739890

Country of ref document: EP

Effective date: 20260113

ENP Entry into the national phase

Ref document number: 2024739890

Country of ref document: EP

Effective date: 20260113

ENP Entry into the national phase

Ref document number: 2024739890

Country of ref document: EP

Effective date: 20260113

ENP Entry into the national phase

Ref document number: 2024739890

Country of ref document: EP

Effective date: 20260113

WWP Wipo information: published in national office

Ref document number: 2024739890

Country of ref document: EP