Method of gene editing using base editors for transgene insertion and multiplex gene editing

The non-viral gene editing method using base editors with single-strand breaks and homology-directed repair addresses genotoxicity and cost issues in CAR T cell therapies, providing precise and efficient transgene insertion and multiplex editing for therapeutic cells.

WO2026082970A1PCT designated stage Publication Date: 2026-04-23CHARITE UNIVSMEDIZIN BERLIN KORPERSCHAFT DES OFFENTLICHEN RECHTS
View PDF 2 Cites 0 Cited by

Patent Information

Authority / Receiving Office
WO · WO
Patent Type
Applications
Current Assignee / Owner
CHARITE UNIVSMEDIZIN BERLIN KORPERSCHAFT DES OFFENTLICHEN RECHTS
Filing Date
2025-10-17
Publication Date
2026-04-23

AI Technical Summary

Technical Problem

Current CAR T cell therapies face challenges such as genotoxicity from random viral integration, high production costs, complexity of multiple genetic modifications, and low editing efficiency, along with risks of off-target effects and treatment failure due to unintended gene transfer.

Method used

A non-viral gene editing method using base editors that introduce single-strand breaks and combine with homology-directed repair to achieve precise and efficient transgene insertion and multiplex gene editing, minimizing genomic instability and off-target effects.

Benefits of technology

This approach reduces genotoxicity, lowers production costs, and enhances editing efficiency, enabling safer and more cost-effective therapies for complex diseases by allowing simultaneous genetic alterations in therapeutic cells.

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure IMGF000094_0001
    Figure IMGF000094_0001
  • Figure 00000120_0000
    Figure 00000120_0000
  • Figure 00000121_0000
    Figure 00000121_0000
Patent Text Reader

Abstract

The invention relates to an in vitro method for modifying double stranded DNA (dsDNA) in a eukaryotic cell. The method comprises: a) introducing a base editor system into the cell, said system comprising i) a base editor, comprising an RNA guided DNA nickase linked to a single-stranded DNA nucleobase modifying enzyme, or a nucleic acid encoding said base editor, and ii) at least two guide RNA (sgRNA) molecules, comprising a first sgRNA that hybridizes to a first target sequence in the dsDNA, and a second sgRNA that hybridizes to a second target sequence in an opposing strand of the dsDNA, b) generating at least two single-strand breaks (SSBs), in opposing strands, of the dsDNA, wherein a first base editor complex, comprising the first sgRNA, introduces a nick cleavage of one strand of the dsDNA in the first target sequence, and a second base editor complex, comprising the second sgRNA, introduces a nick cleavage of the opposite strand of the dsDNA in the second target sequence, thereby producing a 5' or 3' overhang, c) inserting an exogenous DNA repair template sequence in proximity to the two single strand breaks (SSBs). The invention further relates to a genetically modified eukaryotic cell obtained by the method according to the present invention. The invention also relates to a kit for use in the in vitro method.
Need to check novelty before this filing date? Find Prior Art

Description

[0001] METHOD OF GENE EDITING USING BASE EDITORS FOR TRANSGENE INSERTION AND MULTIPLEX GENE EDITING

[0002] DESCRIPTION

[0003] The invention relates to the field of molecular and cell biology, in particular gene editing and the production and use of therapeutic biological cells, such as cellular immunotherapies.

[0004] The invention relates to an in vitro method for modifying double stranded DNA (dsDNA) in a eukaryotic cell, the method comprising: a) introducing a base editor system into the cell, said system comprising i) a base editor, comprising an RNA guided DNA nickase linked to a single-stranded DNA nucleobase modifying enzyme, or a nucleic acid encoding said base editor, and ii) at least two guide RNA (sgRNA) molecules, comprising a first sgRNA that hybridizes to a first target sequence in the dsDNA, and a second sgRNA that hybridizes to a second target sequence in an opposing strand of the dsDNA, b) generating at least two single-strand breaks (SSBs), in opposing strands, of the dsDNA, wherein a first base editor complex, comprising the first sgRNA, introduces a nick cleavage of one strand of the dsDNA in the first target sequence, and a second base editor complex, comprising the second sgRNA, introduces a nick cleavage of the opposite strand of the dsDNA in the second target sequence, thereby producing a 5’ or 3’ overhang, c) inserting an exogenous DNA repair template sequence in proximity to the two single strand breaks (SSB).

[0005] The invention further relates to a genetically modified eukaryotic cell obtained by the method according to the present invention.

[0006] The invention also relates to a kit for use in the in vitro method for modifying double stranded DNA (dsDNA) according to the present invention, comprising a) a base editor, comprising an RNA guided DNA nickase linked to a single-stranded DNA nucleobase modifying enzyme, or a nucleic acid encoding said base editor, and at least two guide RNA (sgRNA) sequences, comprising a first sgRNA that hybridizes to a first target sequence in the dsDNA, and a second sgRNA that hybridizes to a second target sequence in the dsDNA, wherein a first base editor complex, comprising the first sgRNA, is configured to introduce a nick cleavage of one strand of the dsDNA in the first target sequence, and a second base editor complex, comprising the second sgRNA, is configured to introduce a nick cleavage of the opposite strand of the dsDNA in the second target sequence, thereby producing a 5’ or 3’ overhang, and b) an exogenous DNA repair template that hybridizes to the dsDNA in proximity to one or both single strand breaks (SSB), and c) optionally one or more DNA repair modulators.

[0007] BACKGROUND OF THE INVENTION

[0008] Immunotherapies hold significant promise as innovative treatment strategies for cancer and other immune-mediated diseases. Among these, chimeric antigen receptor (CAR) T cell therapies have demonstrated efficacy, particularly in treating hematological malignancies. Since the first CAR T cell therapy was approved in Europe in 2018, six CAR T cell products have received marketing authorization. These approved therapies are autologous, meaning that T cells are collected from the patient's blood, and a new DNA construct encoding a CAR is introduced into these cells ex vivo. This receptor enables the modified T cells to recognize and eliminate cancer cells more effectively.

[0009] In these approved CAR T cell therapies, gene transfer is performed using viral vectors, which randomly integrate the CAR-encoding DNA into the T cell genome. However, this random integration poses a risk of insertional mutagenesis, potentially leading to the transformation of T cells into cancerous cells. For example, 22 cases of T cell lymphoma were reported following the administration of virally produced CAR T cells. It is important to note that this is a rare side effect, given that over 27,000 patients in the United States have undergone CAR T cell therapy (Verdun and Marks, 2024). Additionally, unintended viral gene transfer into malignant blood cancer cells, rather than T cells, can result in treatment failure (Ruella et al., 2018). Beyond these risks, the clinical application of therapies using retroviral or lentiviral vectors in GMP (Good Manufacturing Practice) quality is associated with substantial cost and complicated manufacturing.

[0010] Non-viral gene editing technologies present a promising solution to improve both the safety and cost-effectiveness of CAR T cell therapies. One such alternative is the CRISPR-Cas system, which allows precise gene editing using a simple single guide RNA (sgRNA) to direct the Cas nuclease to a specific DNA sequence within the genome. Upon introducing a double-strand break (DSB) in the DNA, the cell's natural repair mechanisms, such as the non-homologous end joining (NHEJ) pathway and homology-directed repair (HDR), can be harnessed to introduce small insertions or deletions (indels), which can be used for targeted gene knock-out (KO) or to precisely integrate new genetic material, such as a CAR transgene. By supplying an exogenous template, HDR can be harnessed for targeted transgene integration, commonly referred to as knock-in (KI). Leveraging these major DNA repair mechanisms enables precise engineering of cell therapies. This targeted approach further eliminates the need for exogenous promoters, which carry the risk of gene dysregulation and mutagenesis. The precise targeting of defined DNA sequences by CRISPR- Cas provides a significant safety advantage over random viral integration, as off-target effects can be predicted and minimized.

[0011] The recent approval of the CRISPR-Cas-based therapy "Exa-cel" for treating sickle cell disease and beta-thalassemia marks a critical milestone, demonstrating CRISPR-Cas's potential not only for gene transfer but also for targeted gene knockout, such as the disruption of genes or regulatory elements.

[0012] While CAR T cell therapies have demonstrated significant success in treating various blood cancers, their application in other disease areas is still in the early stages. However, there are promising avenues for utilizing gene-edited T cells in the treatment of autoimmune diseases, solid tumors, and organ transplantation. The complexity of these conditions often necessitates multiple genetic modifications within the same T cells. For example, combining viral gene transfer of a CAR with multiple gene knockouts (KOs) can produce T cells that are resistant to lymph-depleting antibodies and immune checkpoints, enabling their use against T cell lymphomas (Georgiadis et al., 2021 ; Diorio et al., 2022).

[0013] CRISPR-Cas-assisted HDR allows the non-viral generation of chimeric antigen receptor (CAR) T cells with controlled expression, for instance by integrating the CAR construct into defined genomic loci such as the T cell receptor alpha constant (TRAC) or CD3£ gene, thereby eliminating the need for exogenous promoters (Muller et al., 2022; Roth et al., 2018; Kath et al., 2024 and Kath et al., 2022). Compared to virally transduced CAR T cells, characterized by random integration and exogenous promoter-driven expression, TRAC- and CD3£-integrated CAR T cells display reduced exhaustion and improved functionality (Kath et al., 2024 and Eyquem et al., 2017). Moreover, CRISPR knock-in at these loci directly disrupts the T cell receptor (TCR) complex, preventing graft- versus-host disease (GvHD) and enabling the development of allogeneic CAR T cell therapies (Qasim et al, 2017 and Torikai et al, 2012). While non-virally engineered allogeneic CAR T cells promise scalable, off-the-shelf therapies derived from healthy donors at reduced cost (Wagner et al., 2022), they must be further engineered to evade rejection by the host’s immune system. This can be achieved by disrupting the genes beta-2-microglobulin (B2M) and class II major histocompatibility complex transactivator (Cl IT A) to abrogate HLA class I and II expression, respectively, and by introducing inhibitory receptors such as HLA-E to suppress NK cell-mediated rejection (Wang et al., 2015; Scharer et al., 2015; McCallion et al. and 2023).

[0014] Allogeneic CAR T cells may benefit from additional gene KOs intended to enhance safety and efficacy, including previously described strategies to prevent T cell exhaustion (Cherkassky et al., 2020) and improve function in the immunosuppressive tumor microenvironment (Tang et al., 2020). Gene editing can also render CAR T cells resistant to immunosuppressive drugs; for example, KO of FK506-binding protein 12 (FKBP12) confers resistance to Tacrolimus (Amini et al., 2020). Such resistance could enable therapy in patients requiring systemic immunosuppression, and, in a different context, allow the use of immunosuppression to limit product rejection, thereby extending the persistence of allogeneic CAR T cells without compromising efficacy. Beyond these rational KO approaches, genome-wide CRISPR KO screens have been employed to uncover genes whose deletion can overcome CAR T cell dysfunction and improve long-term functionality (Wei et al., 2019; Trefny et al., 2023; Freitas et al., 2022 and Shifrut et al, 2018). One of the hits, KO of RAS p21 protein activator 2 (RASA2), exhibited enhanced persistence and effector functions (Carnevale et al., 2022). Although promising on their own, these edits gain full relevance in combination, especially for allogeneic CAR T cells, where multiplex engineering is likely essential for durable safety and efficacy.

[0015] Despite the promise of multiplex gene editing, the introduction of DSBs carries a significant risk of genotoxicity, potentially creating structural variants (SVs) such as large deletions (Alanis-Lobato et al., 2021), chromosome truncations (Lazar et al., 2024), chromosome loss (Nahmad et al., 2022) and translocations, at both on-target and off-target sites (Qasim et al., 2017). These genomic alterations can result in unintended gene fusions, with recurrent balanced chromosomal rearrangements recognized as early initiating events in oncogenesis (Mitelman et al., 2007).

[0016] Thus, while CRISPR-Cas has already entered clinical practice, the next generation of gene editing technologies is rapidly advancing. Among these, "base editing" allows for the precise modification of individual nucleotide bases within a defined DNA region. Base editing employs deaminase enzymes guided by a modified Cas enzyme to the target site in the genome. Unlike traditional CRISPR-Cas, the Cas enzyme in base editing, known as "Cas-nickase," induces a single-strand break ("nick") instead of a double-strand break, reducing the risks associated with DSB-induced toxicity. Adenine base editors (ABE) and cytosine base editors (CBE) use a Streptococcus pyogenes Cas9 (SpCas9) nickase (nCas9) fused to a deaminase to introduce targeted base modifications (Komor et al., 2016 and Gaudelli et al., 2017), with CBE requiring an additional uracil glycosylase inhibitor (UGI) to prevent reversion to the original sequence by base excision repair (Komor et al., 2016). Base editing can introduce premature stop codons or disrupt splice sites for efficient gene KO, supporting its utility in multiplex-editing strategies (Webber et al., 2019; Kluesner et al., 2021 and Kuscu et al., 2017).

[0017] WO2022 / 159753A1 discloses a method utilizing a base editor, at least two primary guide RNAs (gRNAs) for creating one nick on each strand of a dsDNA and a DNA donor template for generating cells with a targeted knock in via HDR. However, this editing system showed low knock in efficiency. When additionally including a gRNA that binds to the sequence of the dsDNA edited by the base editor (retargeting gRNA) knock in efficiency was increased. Further viral donor templates such as AAV are used for targeted knock in.

[0018] Lau et al. (2020) discloses different CRISPR-based strategies for transgene knock in and correction including HDR and base editing. It is further disclosed that HDR can be enhanced by using small molecules such as molecules that modulate HDR pathways or that inhibit NHEJ activity.

[0019] In summary, current CAR T cell therapies face several challenges, including the risk of genotoxicity from the random integration of genetic material using viral vectors, which can lead to mutations and the transformation of T cells into cancerous cells. High production costs associated with generating viral vectors in GMP quality, the complexity and expense of performing multiple genetic modifications and the low efficiency of several editing approaches further complicate these therapies. Additional concerns include the risk of off-target effects, the use of exogenous promoters with mutagenic potential, and the possibility of treatment failure due to unintended gene transfer into non-target cells. While CAR T cells have proven effective in treating blood cancers, their application to autoimmune diseases and solid tumors remains underdeveloped.

[0020] Thus, there is an unmet need for developing efficient non-viral gene-editing strategies that are safe and cost-effective, while showing high editing efficiency, minimizing genotoxicity and avoiding the use of viral vectors.

[0021] SUMMARY OF THE INVENTION

[0022] In light of the prior art, the technical problem underlying the invention was providing alternative or improved in vitro means for modifying the genome of a eukaryotic cell, e.g., for the development of CAR T cells. The present invention seeks to provide such means while avoiding the disadvantages known in the prior art.

[0023] A further objective of the invention was to develop alternative or improved means for minimizing genotoxicity, particularly to mitigate the risks associated with double-strand breaks and random integration of genetic material.

[0024] Another objective of the invention was to develop alternative or improved strategies for non-viral gene editing in therapeutic cells.

[0025] A further objective of the invention was to enable complex genetic modifications, including the ability to perform simultaneous genetic alterations in therapeutic cells, which are employed in developing therapies targeting complex diseases such as autoimmune disorders, organ transplantation, and solid tumors. A further objective of the invention was to enhance the means of production of therapeutic cells, by providing alternative or improved strategies for safer, more cost-effective gene-editing methods, for making multiple genetic alterations to a therapeutic cell.

[0026] These problems are solved by the features of the independent claims. Preferred embodiments of the present invention are provided by the dependent claims.

[0027] The present invention, in one aspect, relates to an in vitro method for modifying double stranded DNA (dsDNA) in a eukaryotic cell, the method comprising: a) introducing a base editor system into the cell, said system comprising a) a base editor, comprising an RNA guided DNA nickase linked to a single-stranded

[0028] DNA nucleobase modifying enzyme, or a nucleic acid encoding said base editor, and b) at least two guide RNA (sgRNA) molecules, comprising a first sgRNA that hybridizes to a first target sequence in the dsDNA, and a second sgRNA that hybridizes to a second target sequence in an opposing strand of the dsDNA, b) generating at least two single-strand breaks (SSBs), in opposing strands, of the dsDNA, a) wherein a first base editor complex, comprising the first sgRNA, introduces a nick cleavage of one strand of the dsDNA in the first target sequence, and a second base editor complex, comprising the second sgRNA, introduces a nick cleavage of the opposite strand of the dsDNA in the second target sequence, thereby producing a 5’ or 3’ overhang, c) inserting an exogenous DNA repair template sequence in proximity to the two single strand breaks (SSBs).

[0029] The present invention is based on the surprising finding that base editing enzymes may facilitate both transgene knock-in and multiplex gene editing within a single transfection process, achieving efficiencies comparable to or exceeding those of conventional CRISPR-Cas9 methods.

[0030] A notable aspect of the invention is its reliance on a non-viral strategy for gene editing. By eliminating the need for viral vectors, this approach significantly reduces genotoxicity by inducing single-strand breaks rather than double-strand breaks, thereby minimizing genomic instability. Additionally, avoiding viral vectors mitigates risks such as insertional mutagenesis, chromosomal translocations, and the costs and risks associated with viral vector production and use.

[0031] One aspect of the present invention that represents significant progress over the prior art lies in the strategic combination of base editing with homology-directed repair, which allows for precise and efficient gene editing without the use of viral delivery systems. This approach offers a safer, cost- effective, and highly efficient alternative to traditional gene editing technologies, marking a significant advancement in the field of genetic engineering.

[0032] Using viral vectors for gene transfer can result in the random integration of genetic material into the genome, posing a risk of insertional mutagenesis. This random integration can potentially lead to the transformation of T cells into cancerous cells, as evidenced by documented cases of patients developing T cell lymphomas after receiving CAR T cell therapy. In contrast, base editing presents several advantages over conventional genome editing techniques. Base editors do not induce double-strand breaks, which significantly reduces the risk of generating off-target indel that can occur during the error-prone non-homologous end-joining (NHEJ) repair process. Additionally, base editors function independently of the cell cycle, making them effective in non-dividing cells, which represents a notable limitation of methods that rely on homology-directed repair (HDR).

[0033] In one embodiment the DNA repair template is a non-viral template.

[0034] In one embodiment the base editor system is introduced without a viral vector.

[0035] In one embodiment the DNA repair template is integrated without a viral vector.

[0036] In embodiments, the method of the present invention comprises generating at least two single-strand breaks (SSBs) in opposing strands of the dsDNA.

[0037] In embodiments, the method of the present invention comprises generating at least two single-strand breaks (SSBs) in opposing strands of the dsDNA, wherein a first base editor complex, comprising the first sgRNA, introduces a nick cleavage of one strand of the dsDNA in the first target sequence and a second base editor complex, comprising the second sgRNA, introduces a nick cleavage of the opposite strand of the dsDNA in the second target sequence, thereby producing a 5’ or 3’ overhang.

[0038] In the context of the present invention, the term "overhang" refers to the single-stranded DNA portion that would be produced if the double-stranded DNA fragments, targeted for modification, were separated following the creation of two single-strand breaks. Each resulting dsDNA fragment would possess an overhang, the length of which corresponds to a spacer region between the two SSBs.

[0039] The nature of the overhang, whether a 3' or 5' on each end of the resulting dsDNA fragment, depends on the positioning of the guide RNAs responsible for directing the guide RNA / DNA nickase complex to the respective strands. Specifically, the relative alignment of the guide RNAs determines whether the breaks occur so that a 3' or 5' overhang is produced, thereby influencing the structural configuration of the modified DNA fragment and facilitating homology-directed repair.

[0040] In embodiments, the method of the invention induces fewer unwanted on- and / or off-target effects, such as micro-insertions and / or deletions (Indels) and / or off-target mutations.

[0041] On- and off-target effects can be measured by sequencing-based techniques with subsequent sequence analysis and comparison, as known to a person skilled in the art. For example, on-Zoff- target effects can be determined by techniques like GUIDE-seq, CAST-Seq or LAM-HTGTS.

[0042] On-target mutations (in particular INDELs) are determined by FOR, following FOR amplicons sequenced by Sanger or next-generation sequencing. Frequencies of on-target mutations are analyzed and compared with reference sequences (wild-type). Genome-wide off-target effects induced by CRISPR / Cas are determined by deep sequencing techniques such as GUIDE-seq, CAST-Seq, AAV-seq, and LAM-HTGTS. Quantification of known potential on- and / or off-target effects such as chromosomal translocations can furthermore be quantified by droplet digital PCR (ddPCR). It is a surprising finding that the method described herein significantly reduces the unwanted on- and off-target gene editing outcomes in the genomic DNA. This surprising finding renders the method of the invention a robust and reliable method for gene correction.

[0043] Surprisingly, high-fidelity gene editing in eukaryotic cells has been demonstrated to be effective through single-strand break-mediated homology-directed repair. This method involves the introduction of base editor system comprising a guide RNA-guided DNA nickase linked to a singlestranded DNA nucleobase modifying enzyme into the target cell, in conjunction with at least two guide RNAs and an exogenous DNA donor repair template. The process results in the creation of at least two single-strand breaks within the double-stranded DNA, forming either 5' or 3' overhangs in the region designated for modification. These single-strand breaks, strategically positioned by the base editor complex, promote the precise incorporation of the exogenous DNA template through homology-directed repair mechanisms, enabling highly accurate gene editing outcomes.

[0044] A key factor contributing to the high specificity of sequence replacement in the method disclosed by the present invention, with minimal off-target or on-target mutations, is the use of a base editor system. Unlike nucleases that generate double-strand breaks with blunt ends, nickases introduce a single-strand break (nick) in one strand of the double-stranded DNA. This controlled nicking reduces the likelihood of unintended genomic alterations and promotes more accurate homology-directed repair in the targeted region. By avoiding the more disruptive double-strand breaks, the method ensures enhanced precision in sequence modification, thereby minimizing genomic instability and off-target effects.

[0045] Preferably, the DNA nickase of the invention is a Cas9 nickase, such as a Cas9 nickase derived from Streptococcus pyogenes Cas9 (SpCas9).

[0046] Following the present invention, the design of the two guide RNAs is specifically configured to ensure that the guide RNA / DNA nickase complexes associated with each respective guide RNA are positioned on opposing strands of the double-stranded DNA targeted for modification. Specifically, the first guide base editor system introduces a nick on the plus strand, while the second guide base editor system introduces a nick on the minus strand. As a result, single-strand nicks are introduced on opposite strands of the dsDNA, leading to the generation of overhangs at each end of the resulting DNA fragment. These complementary overhangs facilitate precise and efficient sequence replacement through homology-directed repair, promoting targeted gene modification with high fidelity and minimal off-target effects.

[0047] In the context of the present invention, the different components of the method, namely a base editor comprising a guide RNA-guided DNA nickase or a nucleic acid molecule encoding a guide RNA- guided DNA nickase, a first guide RNA, a second guide RNA, and an exogenous DNA donor template (comprising a DNA template sequence), may be added sequentially to a cell or a cell culture comprising the cell, or at the same time, for example using a premixed stock solution comprising all or some of the components.

[0048] As used herein, it is understood that replacing a DNA sequence of the dsDNA, positioned in proximity to the two single-strand breaks (SSBs) means replacing a DNA sequence of the dsDNA positioned in proximity to, adjacent to, and / or between the two single-strand breaks (SSBs). The sequence to be replaced in the context of the present invention can be located between the SSBs induced by the nickase, but it can also span the location of one or both SSBs and, therefore, extend beyond the spacer sequence. It is understood that a DNA sequence positioned in proximity to the SSBs comprises any sequence located between the two SSBs, including the entire spacer sequence, but can also be a sequence extending beyond the spacer on one or both sides.

[0049] It is understood that a sequence “adjacent” to an SSB is a sequence that starts with a nucleotide directly next to the SSB / cutting site of the nickase. For example, the spacer sequence located between the two SSBs is on each end “adjacent” to one of the SBBs.

[0050] However, a sequence that is “in proximity” to the SSBs can also comprise sequences whose distal ends start up to 1500 bp outside the spacer sequence, meaning that the distal end of the sequence (as regarded from the next SSB induced by the nickase in the context of the method of the invention) starts in a distance of 1500 bp from the SSB. Accordingly, in embodiments, the sequence to be replaced can span the spacer sequence and up to 1500 bp outside from each of the two SBBs. Preferably, a sequence that is “in proximity” to the SSBs can also comprise sequences whose distal ends start about 1400, 1300, 1200, 1100, 1000, 950, 900, 850, 800, 750, 700, 650, 600, 550, 500, 450, 400, 350, 300, 250, 200, 180, 150, 100, 50, 40, 30, 20, 10, 5, 4, 3, 2, 1 , 0 (= adjacent to the SSB) bp outside the spacer sequence. Furthermore, sequences that are in proximity to the SSBs comprise all sequences that consist of or are comprised by the spacer region.

[0051] In embodiments, the exogenous DNA repair template sequence encodes a recombinant antigenspecific targeting construct, such as a chimeric antigen receptor (CAR), a T cell receptor (TCR), or a construct configured to enhance the persistence and / or potency of cells for therapeutic effect.

[0052] In one embodiment, the exogenous DNA repair template sequence is a homology-directed repair (HDR) donor template sequence.

[0053] In embodiments, the exogenous DNA repair template sequence encodes a chimeric antigen receptor (CAR), preferably configured for expression in a CAR T cell or a CAR NK cell.

[0054] In embodiments, the exogenous DNA repair template sequence encodes a T cell receptor.

[0055] In embodiments, the exogenous DNA repair template sequence encodes a construct configured to enhance the persistence and / or potency of a cell for a therapeutic effect.

[0056] In embodiments, the exogenous DNA repair template sequence encodes an HLA class I histocompatibility antigen, alpha chain E (HLA-E). In embodiments, the exogenous DNA repair template sequence encodes an HLA class I histocompatibility antigen, alpha chain E (HLA-E) to protect a cell from allogeneic NK cell lysis.

[0057] In embodiments, the exogenous DNA repair template sequence encodes a cytokine. In embodiments, the exogenous DNA repair template sequence encodes a cytokine for improved antitumor therapy.

[0058] Advantageously, the in vitro method of the present invention allows for precise targeting and / or insertion of therapeutic constructs such as CARs, T cell receptors, or other constructs configured to enhance the persistence and / or potency of cells, enhancing the specificity and efficacy of the geneediting process, e.g., compared to viral methods. In embodiments, the eukaryotic cell is an effector cell, such as a cytotoxic immune effector cell, preferably a T cell or NK cell. In embodiments, the eukaryotic cell is an effector cell, such as a cytotoxic immune effector cell, preferably a Treg cell.

[0059] In another embodiment, the eukaryotic effector cell is an allogeneic eukaryotic cell with respect to a patient. In one embodiment, the effector cell is an allogeneic effector cell with respect to a patient. In one embodiment, the cytotoxic immune effector cell is an allogeneic cytotoxic immune effector cell with respect to a patient. In one embodiment, the immune effector cell is an allogeneic T cell or an allogeneic NK cell with respect to a patient.

[0060] In another embodiment, the eukaryotic effector cell is an autologous eukaryotic cell with respect to a patient. In one embodiment, the effector cell is an autologous effector cell with respect to a patient. In one embodiment, the cytotoxic immune effector cell is an autologous cytotoxic immune effector cell with respect to a patient. In one embodiment, the immune effector cell is an autologous T cell or an autologous NK cell with respect to a patient.

[0061] In some embodiments, cells can be obtained from a subject different from the intended patient (allogeneic cells). As used herein, a cell is "allogeneic" with respect to a subject if it or one of its progenitor cells is derived from another subject of the same species. As used herein, a cell is "autologous" with respect to a subject if it or its progenitor cells originate from the same subject.

[0062] The method disclosed in the present invention is advantageously applicable to both autologous and allogeneic cells. Notably, the use of allogeneic cells presents a particularly advantageous approach due to their potential for developing "off-the-shelf' therapeutic products, such as generating CAR T cells. In this context, allogeneic cells offer distinct advantages, including the ability to facilitate large- scale production and pre-emptive storage for immediate clinical deployment.

[0063] For the safety of allogeneic therapies, the knockout of the endogenous T cell receptor (TCR) is critical, as it mitigates the risk of graft-versus-host disease (GvHD). GvHD arises when donor- derived CAR-modified T cells, through their endogenous TCR, recognize and attack the recipient's tissues as foreign. The knockout of the endogenous TCR in the allogeneic T cells substantially reduces the risk of GvHD, enhancing the safety profile of such therapies.

[0064] The provision of allogeneic cells, as described in the present invention, enables a more rapid and efficient therapeutic response. This is because allogeneic cells can be sourced and prepared in advance, allowing for quicker availability compared to autologous cells, which must be harvested from the patient. Medical pre-treatments of the patients can result in manufacturing failure or cause autologous cells to be functionally limited. Therefore, selection of allogeneic donor cells with an ideal cellular phenotype can enable optimal performance of the therapy. Allogeneic cells can also be engineered to target various tumor-associated antigens, thus offering greater versatility in treating different types of cancer. Furthermore, manufacturing and storing allogeneic CAR T cells for future use enhances the scalability and accessibility of CAR T cell therapies, potentially broadening their clinical application and making them more readily available to a broader patient population.

[0065] In embodiments, the eukaryotic cell is a T or NK cell and the first and second target sequences of the dsDNA are in an endogenous genomic dsDNA encoding a T cell receptor component, such as in the TRAC locus, CD3C, locus or CD3s locus, or a major histocompatibility complex component, such as in the B2M locus. In one embodiment, the exogenous DNA sequence integrated into the TRAC locus encodes a chimeric antigen receptor (CAR), resulting in a disruption of a T cell receptor (TCR) / CD3 complex expression and integration of the CAR-encoding sequence.

[0066] In one embodiment, the exogenous DNA sequence integrated into the CD3£ locus encodes a truncated chimeric antigen receptor (truncCAR), resulting in disruption of the CD3£ component of the T cell receptor (TCR)ZCDS complex expression, and integration of the truncated chimeric antigen receptor (truncCAR).

[0067] In embodiments, the truncated CAR construct is a CDS zeta-deficient chimeric antigen receptor (CAR) fragment (transgene) that is then integrated into an endogenous CDS zeta / CD247 gene for generating a gene fusion and therefore a functional CAR comprising an exogenous CAR fragment fused with an endogenous CDS zeta domain. The technology employing integrating CDS zeta deficient CAR constructs into endogenous CDS zeta genes is described in detail in WO2022 / 136551.

[0068] In embodiments, the truncated CAR construct comprises three nucleic acid sequence regions, wherein a first nucleic acid sequence region encodes a CD19 chimeric antigen receptor (CAR) fragment without a sequence encoding a functional intracellular domain of a CDS zeta protein, a second nucleic acid sequence region comprises a targeting sequence configured for integrating the construct at an endogenous CDS zeta gene of a human cell, and a third nucleic acid sequence region encoding and or comprising a protein separation site upstream of the first sequence region.

[0069] In embodiments, the targeting sequence of the second nucleic acid sequence region is configured for integrating the construct in-frame with an endogenous CDS zeta encoding sequence, preferably into an exon of an endogenous CDS zeta gene. In embodiments, the protein separation site of the third nucleic acid sequence encodes a first self-cleavage peptide, protease cleavage site or an internal ribosomal entry site (IRES).

[0070] In embodiments, the target site of the CDS zeta gene, into which the first sequence region is integrated, is an exon and / or an intron, preferably exon 2, and is positioned upstream of an endogenous CDS zeta sequence encoding an intracellular domain or fragment thereof, said intracellular domain or fragment thereof comprising one or more immunoreceptor tyrosine-based activation motifs (ITAMs).

[0071] In embodiments, the CAR, when integrated as a truncated CAR into the CDS zeta locus, as provided by the invention, can be expressed under the control of the endogenous CDS zeta promoter.

[0072] In one embodiment, the exogenous DNA sequence integrated into the CD3E locus encodes a chimeric antigen receptor (CAR) or other therapeutic construct, resulting in the disruption of the CDSE component of the T cell receptor (TCR) / CD3 complex expression, and integration of the CAR- encoding sequence.

[0073] In one embodiment, the DNA template for CDSe encodes an antigen binding domain (such as an antibody fragment - scFv) resulting in the incorporation of the CDSe-binding domain into the TCR. Thus, the CDSe component and TCR are not disrupted. In one embodiment, the exogenous DNA sequence integrated into the B2M locus encodes an HLA- E fusion protein, resulting in the disruption of the major histocompatibility complex (MHC) class I molecule expression.

[0074] The method of the present invention advantageously provides a targeted approach by directing gene editing to specific loci within the genome. This precise targeting results in more consistent expression and stable functionality of the inserted genes, in contrast to random integration methods, which can lead to variable gene expression and unpredictable outcomes. By ensuring that transgenes are integrated at predetermined sites, the invention enhances control over gene expression and reduces the risk of insertional mutagenesis. Furthermore, targeted insertion into an open reading frame (ORF) eliminates the need for a promoter, further enhancing the approach's safety.

[0075] In embodiments, the exogenous DNA repair template sequence comprises at least one sequence (homology arm) of at least 45 bp, preferably 400 bp, with at least 90%, preferably 95%, more preferably 100%, sequence identity to a region of the dsDNA sequence, wherein said homology arm is positioned to hybridize to the dsDNA in proximity to the two single strand breaks (SSBs).

[0076] In embodiments, the sequence to be replaced is defined by the design of the DNA template sequence. To induce HDR mediated sequence replacement, the DNA template sequence requires at least one homology arm, which is understood to be a sequence of at least 45, 50, 100, 150, 200, 250, 300, 350, and 400 bp with sequence identity to a region within the DNA sequence sufficient to mediated HDR, such as as at least 90%, preferably 95%, 96%, 97%, 98%, 99%, most preferably 100% sequence identity to a region within the DNA sequence to be replaced. In embodiments, the homology arm has a sequence identity sufficient to enable the hybridization of the homology arm to the corresponding region of the sequence to be replaced. Preferably, the homology arms share this sequence identity with a region located at the end of the sequence to be replaced.

[0077] Preferably, the DNA template has two homology arms, each homologous / sufficiently identical to enable HDR-mediated sequence replacement. In embodiments, the DNA template sequence comprises a sequence unrelated to the replacement sequence located between (flanked by) two homology arms with sufficient sequence identity to a region of the sequence to be replaced to enable HDR-mediated sequence replacement.

[0078] As used herein, the terms sequence identity and sequence homology are used interchangeably.

[0079] Herein, it is understood that a “homologous sequence” has a sufficient sequence identity to a region of the sequence to be replaced to enable replacement through HDR.

[0080] In some embodiment, the DNA template sequence comprises at least one sequence (homology arm) of at least 100 bp, preferably 500 bp, with at least 90%, preferably 95%, more preferably 100%, sequence identity to a region within the DNA sequence to be replaced, wherein said homology arm is positioned to hybridize to the dsDNA molecule to be modified in proximity to the two single-strand breaks (SSBs), preferably wherein the DNA template comprises at least two homology arms with at least 90%, preferably 95%, more preferably 100%, sequence identity to a region within the DNA sequence to be replaced.

[0081] It is a great advantage of the method of the invention that the DNA template sequence can be adjusted as necessary and suited, which can be individually judged by a skilled person depending on the DNA sequence to be replaced and its location within the dsDNA molecule to be modified. There is no strict rule with respect to the designs of the homology arms, which may enable hybridization of a region of the DNA sequence to be replaced, which can be located outside or inside the spacer region or which can comprise the location of one or both SSBs induced in the context of the present invention.

[0082] In some embodiments, the one or two homology arms are positioned to hybridize to the dsDNA molecule to be modified in proximity to, adjacent to, and / or between the two single-strand breaks (SSBs).

[0083] In embodiments, the homology arm includes regulatory elements. In embodiments, the homology arm includes regulatory elements to drive controlled expression of the integrated transgene.

[0084] In embodiments, the method comprises in step a) introducing one or more additional sgRNA sequences, comprising a third sgRNA that hybridizes to a third target sequence in the dsDNA and forming a third base editor complex that introduces a nick cleavage of one strand of the dsDNA and a nucleobase modification in the third target sequence.

[0085] In certain embodiments, the nucleobase modification introduced by the third base editor complex results in a precise alteration of the DNA sequence at the third target site, leading to functional alterations in the gene or regulatory element being modified.

[0086] In embodiments, the third sgRNA targets a splice site to modulate alternative splicing patterns and / or gene expression.

[0087] In embodiments, the third sgRNA results in a precise alteration of the DNA sequence at the third target site from encoding an amino acid to a stop codon to modulate gene expression.

[0088] In embodiments, the third sgRNA introduces a modification to confer resistance to viral infections in the edited cells.

[0089] An advantageous aspect of the present invention is using multiple sgRNA sequences to enable multiplex gene editing. This approach allows for the simultaneous modification of multiple genetic targets, significantly increasing efficiency compared to sequential editing methods. To the best of the inventors' knowledge, no prior study has demonstrated simultaneous editing, such as knock-in (KI) and knock-out (KO), using base editors within a single transfection event. This novel method simplifies the clinical development of T-cell products that require multiple genetic modifications. For example, a single base editor may be employed with multiple guide RNAs targeting different sites in the genome of the therapeutic cell to be modified, thereby enabling multiplex gene editing with a single base editor needing to be transformed into the cells.

[0090] The invention's development of a fully non-viral gene editing strategy for simultaneous knock-in (e.g., a chimeric antigen receptor, CAR) and knock-out of multiple genes is of particular importance. Previous work by the inventors (Glaser et al., 2023) demonstrated the use of different nucleases for knock-in and base editing to prevent chromosomal translocations. However, integrating viral transduction or using multiple nucleases in conjunction with base editing introduces considerable complexity and significantly increases the cost of GMP-grade reagents. While it is possible to clinically translate complex gene-editing strategies, as shown by Chiesa et al. (2023), reducing the complexity of multiplex gene-edited cell products would directly enhance their clinical feasibility and significantly lower associated costs. The present invention addresses this need by simplifying the process, thereby offering a more streamlined, cost-effective approach for clinical applications.

[0091] In embodiments, the third target sequence of the dsDNA is a component of a gene, such as a regulatory sequence, splice donor, splice acceptor, or other sequence associated with expression of said gene, and wherein said modification of the third target sequence leads to disruption of expression of said gene.

[0092] In embodiments, the third target sequence of the dsDNA is located in the open reading frame (ORF) of a gene associated with a functional domain of said gene and wherein said modification of the third target sequence leads to an amino acid change leading to a functional alteration of said gene.

[0093] In embodiments, the third target sequence of the dsDNA is located in the open reading frame (ORF) of a gene wherein said modification of the third target sequence leads to an amino acid change leading to alteration of an epitope. The edited epitope alters the interaction with antigen binding proteins compared to the antigen on non-edited cells. The altered epitope can mask an epitope that was previously detectable on non-edited cells or generate a novel binding site that can be specifically detected.

[0094] In embodiments, the expression of multiple genes is disrupted (multiple gene knockouts) with multiple additional sgRNA sequences, simultaneously with the insertion of an exogenous DNA repair template sequence.

[0095] In embodiments, the method comprises the use of adenine base editors and / or cytosine base editors to achieve targeted nucleobase modifications.

[0096] In embodiments, simultaneous knockout and / or knock-in involves the use of adenine base editors and / or cytosine base editors to achieve targeted nucleobase modifications.

[0097] In embodiments, the disruption of one or more genes involves the use of adenine base editors and / or cytosine base editors. In embodiments, the insertion of an exogenous DNA repair template sequence involves the use of adenine base editors and / or cytosine base editors. In embodiments, the disruption of one or more genes with simultaneous insertion of an exogenous DNA repair template sequence involves the use of adenine base editors and / or cytosine base editors.

[0098] In embodiments, the disruption of one or more genes is achieved through introducing adenine base editors and / or cytosine base editors into the cell. In embodiments, the insertion of an exogenous DNA repair template sequence is achieved is achieved through introducing adenine base editors and / or cytosine base editors into the cell. In embodiments, the disruption of one or more genes with simultaneous insertion of an exogenous DNA repair template sequence is achieved through introducing of adenine base editors and / or cytosine base editors into the cell.

[0099] In embodiments, the disruption of one or more genes using the methods of the present invention leads to lower levels of chromosomal translocation. In embodiments, the insertion of an exogenous DNA repair template sequence using the methods of the present invention leads to reduced chromosomal translocation. In embodiments, the disruption of one or more genes with simultaneous insertion of an exogenous DNA repair template sequence leads to lower levels of chromosomal translocation in comparison to using nuclease-mediated methods. In embodiments, reduced translocations occur with the present invention in comparison to methods using nuclease-mediated gene editing. In general, knock-ins and multiplexed editing are possible with minimal translocations, when using the approach of the present invention.

[0100] In embodiments, one or more exogenous DNA repair template sequences are inserted into the cell.

[0101] In embodiments, one or more exogenous DNA repair template sequences are inserted into the cell in combination with multiple nucleobase modifications.

[0102] In embodiments, one sgRNA guides the base editing machinery to a specific locus for modification, and the second sgRNA directs the insertion of the exogenous template.

[0103] In embodiments, the method of the present invention further comprises the use of specific sgRNA combinations to optimize editing efficiency and minimize off-target effects.

[0104] Selecting appropriate sgRNAs is crucial for creating optimal conditions for intended genetic modifications, such as knock-ins (KIs) or knockouts (KOs). The choice of sgRNAs is informed by the necessity of positioning the editing window effectively, ensuring that the targeted bases are accessible for efficient modification by the base editors. Specifically, minimizing certain nucleobases within the editing window enhances editing efficiency. For instance, reducing adenine bases for Adenine Base Editors (ABEs) or cytosine bases for Cytosine Base Editors (CBEs) can lead to more precise genetic edits. Furthermore, the spacing between DNA nicks created by different sgRNAs is a critical factor that influences the success of the editing process. Maintaining an optimal distance between these nicks facilitates favorable conditions for homology-directed repair, the preferred pathway for the high-fidelity integration of exogenous DNA. The careful selection and combination of sgRNAs, therefore, not only enhance KI and KO efficiencies but also reduce off-target effects, resulting in a higher overall success rate in achieving the desired genetic modifications.

[0105] The method described in the present invention advantageously facilitates a wide array of genetic modifications within a single editing event. By enabling the disruption of gene expression, the introduction of functional alterations, and the simultaneous editing of multiple genes, this invention provides a versatile strategy for comprehensive gene manipulation.

[0106] The capacity for multiplex gene editing reduces the reliance on sequential editing processes, thereby enhancing overall efficiency. This simplification can result in shorter development timelines, reduced costs, and diminished risks of errors or unintended consequences during the editing process.

[0107] Moreover, the ability to execute multiplex editing, along with targeted gene disruption and functional modification, supports the advancement of innovative therapeutic strategies. For example, creating cell therapies that necessitate both the insertion of therapeutic genes and the knockout of inhibitory genes can significantly enhance treatment efficacy for diseases such as cancer and genetic disorders.

[0108] Overall, the present invention offers a flexible and adaptable approach that can be tailored to various cell types and genetic contexts. This adaptability is particularly advantageous for addressing a wide range of diseases and conditions, thereby broadening the applicability of the gene editing platform.

[0109] In embodiments, the gene is selected from FKBP12, NR3C1, PPIA, CD38, CD45, CD52, DNMT3A, Regnase-1, NR4A1, NR4A2, NR4A3, RASA2, PIC3CD, PDCD1, CTLA4, CISH, PTPN2, KLRC1, SOCS1, TIM3, LAGS, CDS, CD7, CD30, SUV39H1, TGFBR2, CD3d, CD3e, CD3g, CD3z, TRBC1, TRBC2, B2M, CIITA, HLA-A, HLA-B, HLA-C, CD54, CD58, CD74, CD16, or a non-coding RNA relevant to T cell function, such as a microRNA or a long non-coding RNA.

[0110] In embodiments, the target gene is a T cell receptor complex component.

[0111] In embodiments, the target gene is a co-receptor modulating T cell activation.

[0112] In embodiments, the target gene is in involved T cell activation.

[0113] In embodiments, the target gene is involved in T cell signaling.

[0114] In embodiments, the target gene is associated with the MHC class I complex. In embodiments, the target gene is an MHC class I molecule. In embodiments, the target gene presents antigens to cytotoxic T cells. In embodiments, the target gene is involved in antigen presentation and / or T cell- APC interactions.

[0115] In embodiments, the target gene is a cytokine.

[0116] In embodiments, the target gene is involved in intracellular signaling. In embodiments, the target gene is involved in protein folding.

[0117] In embodiments, the target gene is involved in immune regulation.

[0118] In embodiments, the single-stranded DNA nucleobase modifying enzyme is an adenosine deaminase, a cytidine deaminase, cytosine-DNA glycolase, thymidine-DNA glycolase, or fusion constructs containing deaminases and glycolases, such as a cytodine deaminase with a uracil glycolase.

[0119] In embodiments, the base editor is a Cytosine Base Editor (CBE) and / or an Adenine Base Editor (ABE) and / or a Glycosylase Base Editor (GBE).

[0120] In embodiments, the base editor is an Adenine Base Editor (ABE). In embodiments, the ABE substitutes adenine (A) nucleotides with guanine (G) nucleotides. In embodiments, the ABE substitutes adenine-threonine (A-T) base pairs with guanine-cytosine (G-C) base pairs.

[0121] In embodiments, the base editor is a Cytosine Base Editor (CBE). In embodiments, the CBE substitutes cytosine (C) nucleotides with threonine (T) nucleotides. In embodiments, the CBE substitutes cytosine-guanine (C-T) base pairs substitutes threonine-adenine (T-A) base pairs.

[0122] In embodiments, the base editor is a Glycosylate Base Editor (GBE). In embodiments, the GBE substitutes guanine (G) nucleotides with adenine (A) nucleotides and / or cytosine (C) nucleotides.

[0123] In one embodiment, the method of the present invention combines several base editors. In embodiments, the base editor is an Adenine Base Editor (ABE) combined with a Cytosine Base Editor (CBE). In embodiments, the base editor is an Adenine Base Editor (ABE) combined with a Glycosylase Base Editor (GBE). In embodiments, the base editor is a Cytosine Base Editor (CBE) combined with a Glycosylase Base Editor (GBE). In embodiments, the base editor is a Cytosine Base Editor (CBE) and an Adenine Base Editor (ABE) and a Glycosylase Base Editor (GBE). In embodiments, the first and / or second sgRNAs, and optionally additional sgRNAs, are configured such that an editing window within the target DNA sequence is located within positions 3-15 of the 20 base pair protospacer region of the sgRNA.

[0124] In embodiments, the first and / or second sgRNAs, and optionally additional sgRNAs, are configured such that the specific target nucleobase within the target DNA sequence is located within the editing window at the positions 3, 4, 5, 6, 7, 8, 9, 10, 11 , 12, 13, 14, 15 of the 20 base pair protospacer region of the sgRNA.

[0125] In embodiments, the editing window is located at positions 3 to 20, preferably 4 to 10, most preferably 4 to 8 of the 20 base pair protospacer region of the sgRNA.

[0126] In embodiments, the editing window overlaps with the target nucleobases within positions 3-15 of the 20 base pair protospacer region of the sgRNA.

[0127] In embodiments, the sgRNA configuration is adjusted based on the specific base composition (adenine or cytosine) within the editing window and available PAMs surrounding the target sequence.

[0128] In embodiments, the sgRNA combinations lack adenine bases in the editing window within the target DNA sequence when using adenine base editors. In embodiments, the editing window lacks adenine bases within the target DNA sequence when using adenine base editors.

[0129] In embodiments, the sgRNA combination comprises 1 , 2, 3, 4, 5, 6, 7, 8 adenine bases in the editing window within the target DNA sequence when using adenine base editors.

[0130] In embodiments, the sgRNA combinations lack cytosine bases in the editing window within the target DNA sequence when using cytosine base editors. In embodiments, the editing window lacks cytosine bases within the target DNA sequence when using cytosine base editors.

[0131] In another embodiment, the editing window aligns with regulatory regions or functional domains of the target gene.

[0132] The present invention advantageously optimizes the design of the editing window to minimize bases that could lead to unintended modifications. For instance, reducing the number of adenines in the target region when employing an adenine base editor (ABE) significantly decreases on-target undesired A-to-G conversions, enhancing specificity and precision. Additionally, the size and placement of the editing window have been strategically optimized to achieve maximum editing efficiency. For example, without being bound by theory, certain sgRNA combinations that promote high knock-in (KI) efficiency exhibited specific characteristics, notably reducing adenines within the editing window when using ABE. Low levels of adenines yield efficient knock ins. For example, up to 4 adenines within the combined editing window of the 2 sgRNAs still yielded efficient KI. Additionally, sgRNAs lacking adenines in this region have demonstrated significantly higher KI frequencies, highlighting the critical role of precise editing window design in achieving optimal gene-editing outcomes.

[0133] A person skilled in the art is able to select an appropriate base editing system and the sgRNA to adjust the position of the editing window without undue effort. In embodiments, the base editor is an ABE, and the first and / or second sgRNA lacks adenine bases in the editing window and / or wherein the base editor is a CBE, and the first and / or second sgRNA lacks cytosine bases in the editing window.

[0134] In embodiments, the base editor is an ABE, and the first and second sgRNA together have a total of less than 6 adenine bases in the editing window and / or wherein the base editor is a CBE, and the first and second sgRNA together have a total of less than 6 cytosine bases in the editing window.

[0135] In embodiments, the base editor is an ABE, and the first and second sgRNA together have a total of less than 6, 5, 4, 3, 2 adenine bases in the editing window.

[0136] In embodiments, the base editor is a CBE, and the first and second sgRNA together have a total of less than 6, 5, 4, 3, 2 cytosine bases in the editing window.

[0137] In one embodiment the sgRNAs used for the insertion of a transgene can introduce base modifications within the editing window of the base editor. In embodiments, first and / or second sgRNA introduce base edits that result in efficient gene knockout at of the targeted gene. In embodiments, the targeted base targets a splice acceptor, splice donor, start codon, regulatory element. In embodiments, the base edit changes a functionally important amino acid in the coding sequence. In embodiments, the base editor introduces a premature stop codon.

[0138] Surprisingly, the present invention demonstrates that minimizing the presence of specific target nucleobases within the editing window significantly enhances both the specificity and efficiency of the base editing process. It was unexpectedly found that a reduced concentration of adenine bases correlates with increased knock-in (KI) rates when employing adenine base editors (ABEs). Similarly, a decrease in cytosine bases leads to enhanced editing efficiency with cytosine base editors (CBEs). Thus, the advantage of the method disclosed in the present invention lies in its ability to achieve higher editing precision and efficiency by constraining the number of target nucleobases within the editing window.

[0139] This method significantly reduces the likelihood of unwanted on-target effects and unintended modifications by strategically configuring sgRNAs to limit the presence of adenines or cytosines in the editing window. This increased precision is critical for applications that necessitate high fidelity and safety in gene editing, particularly in therapeutic contexts.

[0140] This method significantly reduces the likelihood of off-target effects and unintended modifications by utilizing in-silico prediction of potential off target sites during the sgRNAs design process. Furthermore, off-target DSB are reduced when using base editors relying on SSBs compared to nuclease mediated KI strategies. This reduced the risk for genotoxic events, which is particularly crucial for therapeutic applications.

[0141] In embodiments, the method optionally comprises introducing to said cell one or more DNA repair modulators.

[0142] In embodiments, the method comprises introducing to said cell one or more HDR enhancers. Nonlimiting examples of HDR enhancers include Wortmannin, SCR7, L755507, EPZ5676, Rucaparib, Pevonedistat / MLN4924, Brefeldin A, Alt-R HDR Enhancer (V1), XL413, NU7441 , Trichostatin A, CRISPY mix, Romidepsin, Nedisertib / M3814, and Alt-R HDR Enhancer V2. In embodiments, the method comprises contacting said cell with one or more DNA-PK inhibitors and / or DNA PolQ inhibitors. Non-limiting examples of DNA-PK inhibitors are ART558, NU7026, NU7441 , M3814, AZD7648, and KU-0060648.

[0143] Any DNA-PK inhibitor, DNA PolQ inhibitor, or HDR enhancer known in the art is suitable for the method of the present invention. A skilled person is capable of selecting a suitable inhibitor or enhancer without undue effort.

[0144] The introduction of additional DNA repair modulators can enhance homology-directed repair (HDR), thereby improving the integration efficiency of exogenous DNA sequences during knock-in processes. By strategically modifying the DNA repair pathways, it is possible to achieve more efficient and precise gene edits. This modulation increases the likelihood of desired repair outcomes, favoring HDR over non-homologous end joining (NHEJ), often associated with unwanted genetic alterations. Consequently, employing DNA repair modulators can significantly improve the efficacy and specificity of base editing techniques, facilitating the accurate incorporation of therapeutic genes or sequences into targeted genomic loci.

[0145] Various DNA repair modulators are known in the art and a person skilled in the art is able to select a modulator according to the specific need.

[0146] In embodiments, one or more PAM regions on the dsDNA targeted by said sgRNAs are oriented facing outwards.

[0147] In the context of the present invention, the term "facing outwards" refers to the protospacer adjacent motif (PAM) being oriented outward, or external, to the region located between the target sequences of the two guide RNAs used in the method.

[0148] In the context of the present invention, the term "facing outwards" refers to the PAM being oriented outward, or external, to the region located between the target sequences of the two guide RNAs used in the method.

[0149] For example, when utilizing a Cas9, the PAM sequences of the two target sites are positioned outside the spacer region between the two single-strand breaks. This is because the PAM is located at the 3' end of the target sequence, and the nicking occurs three bases upstream of the PAM site.

[0150] Thus, in a preferred embodiment where a "PAM-out" configuration is employed, the guide RNAs are designed such that the PAM sequences are external to the modification region, leading to the generation of 5' overhangs at the ends of the resulting DNA fragments. This strategic configuration optimizes the generation of overhangs and facilitates precise homology-directed repair, contributing to the high fidelity of the gene editing process.

[0151] A person skilled in the art is able to modify these configurations to suit alternative nickases, such as Cas9 H840A or other variants, which may require different guide RNA arrangements or spacer design. The specific positioning of the guide RNAs, the PAM orientation and the resulting overhangs can be adjusted based on the cleavage properties of the particular nickase being used. For instance, different nickases may generate single-strand breaks at alternative positions relative to the PAM, necessitating changes in the guide RNA design and the configuration of the target region to optimize the resulting 5' or 3' overhangs. These modifications ensure efficient and precise gene editing through homology-directed repair in accordance with the method of the invention. In embodiments, the distance between the two single strand breaks on the dsDNA is more than 29 bases.

[0152] In embodiments, the distance between the two single strand breaks on the dsDNA is more than 25 bases, preferably more than 50 bases, more preferably more than 100 bases.

[0153] In embodiments, the distance between the two single strand breaks on the dsDNA is less than 250 bases, preferably less than 200 bases.

[0154] In embodiments, the distance between two single strand breaks on the dsDNA is between 30 and 250 bases, more preferably between 50 and 200 bases. In embodiments, the distance between two single strand breaks on the dsDNA is 30, 40, 50, 60, 70, 80, 90, 100, 110, 120, 130, 140, 150, 160, 170, 180, 190, 200, 210, 220, 230, 240, 250 bases. The distance between two single strand breaks on the dsDNA may also fall within the range formed between any to endpoints mentioned in the list.

[0155] Optimizing the distance between DNA nicks is essential for facilitating effective base editing and the integration of exogenous DNA sequences. The specified distance between two single-strand breaks (SSBs) in double-stranded DNA (dsDNA) is particularly beneficial for achieving efficient homology- directed repair (HDR). Maintaining a maximum distance of less than 250 bases ensures that the repair machinery can effectively bridge the gaps between the breaks. This optimization stabilizes the repair intermediates and aligns with the physical constraints imposed by the cellular repair machinery, enhancing the likelihood of successful repair outcomes. Furthermore, this constraint helps maintain the structural integrity of the repair template and minimizes the potential for erroneous repair processes, thereby improving the overall efficiency and precision of base editing as well as the successful incorporation of therapeutic sequences into the genome.

[0156] In one aspect, the present invention relates to a genetically modified eukaryotic cell obtained by the method according to present invention.

[0157] Advantageously, the method of the present invention produces genetically modified cells featuring one or more precise genetic alterations. These cells are characterized by targeted genetic modifications attained through the precise application of base editing techniques and the strategic utilization of sgRNAs and DNA repair modulators. As a result, the genetically modified eukaryotic cells generated by this method exhibit specific and intended genetic and functional changes, enhancing their potential for therapeutic applications.

[0158] In embodiments, the genetically modified eukaryotic cell comprises at least one insertion of an exogenous DNA donor template sequence and at least one base edit in the dsDNA produced by the base editor system in proximity to said exogenous DNA donor template sequence.

[0159] In embodiments, the genetically modified eukaryotic cell comprises 1 , 2, 3, 4, 5, 6, 7, 8, 9, 10, 11 , 12, 13, 14, 15, 16, 17, 18, 19, 20, 21 , 22, 23, 24, 25, 26, 27, 28, 29, 30 insertions of an exogenous DNA donor template sequence. In embodiments, the genetically modified eukaryotic cell comprises at least one insertion of 1 , 2, 3, 4, 5, 6, 7, 8, 9, 10, 11 , 12, 13, 14, 15, 16, 17, 18, 19, 20, 21 , 22, 23, 24, 25, 26, 27, 28, 29, 30 exogenous DNA donor template sequences.

[0160] In embodiments, the genetically modified eukaryotic cell comprises 1 , 2, 3, 4, 5, 6, 7, 8, 9, 10, 11 ,

[0161] 12, 13, 14, 15, 16, 17, 18, 19, 20, 21 , 22, 23, 24, 25, 26, 27, 28, 29, 30 insertions of 1 , 2, 3, 4, 5, 6, 7, 8, 9, 10, 11 , 12, 13, 14, 15, 16, 17, 18, 19, 20, 21 , 22, 23, 24, 25, 26, 27, 28, 29, 30 exogenous DNA donor template sequences.

[0162] In further embodiments, the genetically modified eukaryotic cell comprises 1 , 2, 3, 4, 5, 6, 7, 8, 9, 10, 11 , 12, 13, 14, 15, 16, 17, 18, 19, 20, 21 , 22, 23, 24, 25, 26, 27, 28, 29, 30 base edits in the dsDNA produced by the base editor system.

[0163] In embodiments, the integration of exogenous DNA sequences comprises the integration of a chimeric antigen receptor (CAR) into a specific genomic locus, e.g., in the TRAC locus.

[0164] The combination of exogenous DNA insertion and concurrent base editing enables the creation of cells with enhanced or novel functionalities, which is one important objective of the described gene editing methods.

[0165] In embodiments, the base editors mediate a knockout of the cell population that did not integrate the DNA template into the target gene.

[0166] The present invention provides a novel method for gene editing that utilizes base editing techniques to achieve efficient knockout (KO) of cell populations that fail to integrate the DNA template into the target gene. An embodiment of this method involves using two sgRNAs in conjunction with base editing to mediate targeted gene modifications. For instance, as shown below, in the case of the p2- microglobulin (B2M) gene, one of the sgRNAs employed in the knockout process was designed to disrupt a critical splice site within the gene. This unexpected outcome not only enhances the overall efficiency of the gene editing process but also broadens the potential applications of this method, providing a significant advantage over traditional gene editing techniques.

[0167] In a further aspect, the invention relates to a kit for use in the in vitro method for modifying double stranded DNA (dsDNA) according to any one of claims 1 to 16, the kit comprising: a) a base editor system, said system comprising i. a base editor, comprising an RNA guided DNA nickase linked to a singlestranded DNA nucleobase modifying enzyme, or a nucleic acid encoding said base editor, and ii. at least two guide RNA (sgRNA) sequences, comprising a first sgRNA that hybridizes to a first target sequence in the dsDNA, and a second sgRNA that hybridizes to a second target sequence in the dsDNA, wherein a first base editor complex, comprising the first sgRNA, is configured to introduce a nick cleavage of one strand of the dsDNA in the first target sequence, and a second base editor complex, comprising the second sgRNA, is configured to introduce a nick cleavage of the opposite strand of the dsDNA in the second target sequence, thereby producing a 5’ or 3’ overhang, and b) an exogenous DNA repair template that hybridizes to the dsDNA in proximity to one or both single strand breaks (SSBs). c) and optionally one or more DNA repair modulators. Further embodiments of the invention employ specific gRNA sequences and / or template repair sequences in order to carry out the method of the present invention.

[0168] The invention therefore also relates to isolated nucleic acid molecules, including for example gRNA or repair templates (also referred to as HDRT (homology directed repair templates), knock-in templates, knock-ins, and the like) as such. These molecules may represent an embodiment of the invention, to be employed in the inventive method, or an independent aspect of the invention directed to the molecules themselves, preferably defined by the sequences presented below.

[0169] Table 1. Preferred nucleotide sequences of the present invention:

[0170] SEQ

[0171] Sequence Description ID No.

[0172] 1 AGAGTCTCTCAGCTGGTACA TRAC sense T 1

[0173] 2 GAGAATCAAAATCGGTGAAT TRAC sense T2

[0174] 3 aaagtcagatttgttgctccagg _ TRAC sense T3

[0175] 4 cttacctgggctggggaagaagg TRAC sense T4

[0176] 5 agctgcccttacctgggctgggg _ TRAC sense T5

[0177] 6 caccaaagctgcccttacctggg TRAC sense T6

[0178] 7 AGCCCAGGTAAGGGCAGCTT TRAC antisense T7

[0179] 8 TGTGCTAGACATGAGGTCTA TRAC antisense T8

[0180] 9 AGAGCAACAGTGCTGTGGCC TRAC antisense T9

[0181] 10 aacaaatgtgtcacaaagta _ TRAC antisense T 10

[0182] 11 gacaccttcttccccagccc _ TRAC antisense T 11

[0183] 12 ttcttccccagcccaggtaaggg _ TRAC antisense T 12

[0184] 13 acatccccattaccccaggg _ CD3z sense Z1

[0185] 14 tctgtgccaagagataaagc _ CD3z sense Z2

[0186] 15 tccagcaggtagcagagttt _ CD3z sense Z3

[0187] 16 gaatgacaccatagatgaag _ CD3z sense Z4

[0188] 17 tgccttgttcctgagagtga _ CD3z antisense Z5

[0189] 18 gatggaatcctcttcatcta _ CD3z antisense Z6

[0190] 19 aggtgggtaccactgggctt _ CD3z antisense Z7

[0191] 20 gagtgaaggtgggtaccact CD3z antisense Z8

[0192] 21 ACAAAGTCACATGGTTCACA B2M sense B1

[0193] 22 CTTACCCCACTTAACTATCT B2M sense B2

[0194] 23 CTTGCCCCACTTAACTATCT B2M sense B3

[0195] 24 TGGGCTGTGACAAAGTCACA B2M sense B4

[0196] 25 AGTCACATGGTTCACACGGC B2M sense B5

[0197] 26 AAGCTGCTTTGATATAAAAA B2M antisense B6

[0198] 27 ATCTGCATATTGGGATTGTC B2M antisense B7

[0199] 28 ACAGCCCAAGATAGTTAAGT B2M antisense B8

[0200] 29 TTTGATATAAAAAAGGTCTA B2M antisense B9

[0201] 30 AGAGTCTCTCAGCTGGTACA TRAC Cas9

[0202] 31 gagtctctcagctggtacacggc TRAC Cas12a

[0203] 32 GCTTACAGATCTTGCCCCGC RASA2 sgRNA BE #1

[0204] 33 ATTTACCTGAACCTCTGAAT RASA2 sgRNA BE #2

[0205] 34 CCCTTACCAGGCTTGATGAG RASA2 sgRNA BE #3

[0206] 35 CTCACCGTCTCCTGGGGAGA FKBP12 sgRNA BE #1

[0207] 36 CTCACCGGTGTAGTGCACCA FKBP12 sgRNA BE #2

[0208] 37 CTACTCACCGTCTCCTGGGG FKBP12 sgRNA BE #3

[0209] TTTCAGGTTTCCTTGAGTGGCAGGCCAGGCCTGGCCGTGAACGTT HDRT1

[0210] 38 CACTGAAATCATGGCCTCTTGGCCAAGATTGATAGCTTGTGCCTGT (T1 / T7) CCCTGAGTCCCAGTCCATCACGAGCAGCTGGTTTCTAAGATGCTAT TTCCCGTATAAAGCATGAGACCGTGACTTGCCAGCCCCACAGAGC CCCGCCCTTGTCCATCACTGGCATCTGGACTCCAGCCTGGGTTGG GGCAAAGAGGGAAATGAGATCATGTCCTAACCCTGATCCTCTTGTC CCACAGATATCCAGAACCCTGACCCTGCCTCCGGATCCGGAGCCA CCAACTTCAGCCTGCTGAAGCAGGCCGGCGACGTGGAAGAAAATC CTGGGCCCATGGCTCTTCCTGTGACTGCCCTTCTGCTGCCCCTGG CTCTCCTGCTCCATGCTGCCCGGCCAGACATCCAGATGACACAGA CAACCAGCAGCCTCAGCGCCAGCCTGGGGGACAGAGTGACCATTA GCTGCCGGGCCTCTCAGGACATCAGCAAATACCTGAACTGGTACC AGCAGAAACCAGATGGCACTGTCAAGCTGCTGATTTACCACACCTC CAGGCTCCACAGCGGCGTGCCCAGTCGCTTCAGCGGCAGTGGGA GCGGGACAGATTATTCCCTCACAATCTCCAACCTGGAGCAGGAAG ATATTGCCACATACTTCTGCCAGCAAGGCAACACCCTGCCATACAC ATTTGGAGGCGGCACCAAATTGGAGATCACCGGCGGTGGTGGATC TGGAGGAGGGGGCAGCGGAGGTGGCGGCTCTGAGGTGAAACTGC AGGAGAGTGGCCCTGGCCTGGTGGCTCCCAGCCAGAGCCTTTCT GTCACCTGCACCGTGTCTGGGGTGTCCCTGCCTGACTATGGAGTC TCCTGGATCCGGCAGCCTCCAAGAAAAGGACTGGAATGGCTGGGC GTCATCTGGGGAAGTGAGACCACCTACTATAATTCAGCCCTCAAGT CCCGGCTCACCATCATTAAGGACAACTCCAAATCCCAGGTGTTCCT GAAGATGAATTCTCTCCAGACTGATGACACAGCCATCTACTACTGT GCCAAGCATTATTATTACGGCGGGTCCTATGCCATGGACTACTGG GGCCAGGGGACCAGTGTCACTGTTTCTTCTGAAAGCAAGTATGGC CCCCCATGCCCTCCCTGCCCCGGACAGCCGCGGGAGCCCCAGGT CTACACCCTCCCACCCTCCAGAGATGAGCTGACCAAGAACCAGGT TTCACTGACATGCCTGGTGAAGGGCTTCTACCCCTCTGACATTGCT GTGGAGTGGGAGAGCAATGGGCAGCCAGAGAACAACTACAAGAC CACACCTCCTGTCCTGGACAGCGACGGCTCCTTCTTCCTCTACAGT AAGCTCACTGTGGACAAGAGCCGCTGGCAGCAGGGCAATGTCTTC TCCTGCAGCGTGATGCACGAGGCCCTGCACAATGCCTATACCCAG AAGTCACTCTCTCTGAGCCCTGGAAAGAAGGATCCCAAGTTTTGGG TCCTCGTGGTGGTCGGGGGTGTGCTGGCCTGCTACAGCTTGCTGG TCACAGTGGCCTTCATCATCTTCTGGGTGCGCTCCAAGAGGAGCC GGCTGCTTCACAGTGATTACATGAACATGACCCCCAGGAGGCCAG GACCCACCAGGAAGCACTACCAGCCCTACGCTCCCCCGCGGGAC TTTGCTGCTTACCGCAGCAGGGTCAAATTTTCTAGATCTGCAGATG CGCCGGCCTATCAAcaaggccagaaccagctcTATAACGAGCTCAATCTA GGACGAAGAGAGGAGTACGATGTTTTGGACAAGAGACGTGGCCG GGACCCTGAGATGGGGGGAAAGCCGAGAAGGAAGAACCCTCAGG AAGGCCTGTACAATGAACTGCAGAAAGATAAGATGGCGGAGGCCT ACAGTGAGATTGGGATGAAAGGCGAGCGCCGGAGGGGCAAGGGG CACGATGGCCTTTACCAGGGTCTCAGTACAGCCACCAAGGACACC TACGACGCCCTTCACATGCAGGCCCTGCCCCCTCGCTAAtaagaattct aactagagctcgctgatcagcctcgactgtgccttctagttgccagccatctgttgtttgcccctccccc gtgccttccttgaccctggaaggtgccactcccactgtcctttcctaataaaatgaggaaattgcatcg cattgtctgagtaggtgtcattctattctggggggtggggtggggcaggacagcaagggggaggat tgggaagagaatagcaggcatgctggggactttggtgccttcgcaggctgtttccttgcttcaggaat ggccaggttctgcccagagctctggtcaatgatgtctaaaactcctctgattggtggtctcggccttatc cattgccaccaaaaccctctttttactaagaaacagtgagccttgttctggcagtccagagaatgac acgggaaaaaagcagatgaagagaaggtggcaggagagggcacgtggcccagcctcagtct ctccaactgagttcctgcctgcctgcctttgctcagactgtttgccccttactgctcttctaggcctcattct aagccccttctccaagttgcctctccttatttctccctgtctgccaaaaaatctttcccagctcactaagt cagtctcacgcagtcactc

[0211] TTTCAGGTTTCCTTGAGTGGCAGGCCAGGCCTGGCCGTGAACGTT CACTGAAATCATGGCCTCTTGGCCAAGATTGATAGCTTGTGCCTGT HDRT2

[0212] 39 CCCTGAGTCCCAGTCCATCACGAGCAGCTGGTTTCTAAGATGCTAT (T1 / T8) TTCCCGTATAAAGCATGAGACCGTGACTTGCCAGCCCCACAGAGC CCCGCCCTTGTCCATCACTGGCATCTGGACTCCAGCCTGGGTTGG GGCAAAGAGGGAAATGAGATCATGTCCTAACCCTGATCCTCTTGTC CCACAGATATCCAGAACCCTGACCCTGCCTCCGGATCCGGAGCCA CCAACTTCAGCCTGCTGAAGCAGGCCGGCGACGTGGAAGAAAATC CTGGGCCCATGGCTCTTCCTGTGACTGCCCTTCTGCTGCCCCTGG CTCTCCTGCTCCATGCTGCCCGGCCAGACATCCAGATGACACAGA CAACCAGCAGCCTCAGCGCCAGCCTGGGGGACAGAGTGACCATTA GCTGCCGGGCCTCTCAGGACATCAGCAAATACCTGAACTGGTACC AGCAGAAACCAGATGGCACTGTCAAGCTGCTGATTTACCACACCTC CAGGCTCCACAGCGGCGTGCCCAGTCGCTTCAGCGGCAGTGGGA GCGGGACAGATTATTCCCTCACAATCTCCAACCTGGAGCAGGAAG ATATTGCCACATACTTCTGCCAGCAAGGCAACACCCTGCCATACAC ATTTGGAGGCGGCACCAAATTGGAGATCACCGGCGGTGGTGGATC TGGAGGAGGGGGCAGCGGAGGTGGCGGCTCTGAGGTGAAACTGC AGGAGAGTGGCCCTGGCCTGGTGGCTCCCAGCCAGAGCCTTTCT GTCACCTGCACCGTGTCTGGGGTGTCCCTGCCTGACTATGGAGTC TCCTGGATCCGGCAGCCTCCAAGAAAAGGACTGGAATGGCTGGGC GTCATCTGGGGAAGTGAGACCACCTACTATAATTCAGCCCTCAAGT CCCGGCTCACCATCATTAAGGACAACTCCAAATCCCAGGTGTTCCT GAAGATGAATTCTCTCCAGACTGATGACACAGCCATCTACTACTGT GCCAAGCATTATTATTACGGCGGGTCCTATGCCATGGACTACTGG GGCCAGGGGACCAGTGTCACTGTTTCTTCTGAAAGCAAGTATGGC CCCCCATGCCCTCCCTGCCCCGGACAGCCGCGGGAGCCCCAGGT CTACACCCTCCCACCCTCCAGAGATGAGCTGACCAAGAACCAGGT TTCACTGACATGCCTGGTGAAGGGCTTCTACCCCTCTGACATTGCT

[0213] GTGGAGTGGGAGAGCAATGGGCAGCCAGAGAACAACTACAAGAC CACACCTCCTGTCCTGGACAGCGACGGCTCCTTCTTCCTCTACAGT AAGCTCACTGTGGACAAGAGCCGCTGGCAGCAGGGCAATGTCTTC TCCTGCAGCGTGATGCACGAGGCCCTGCACAATGCCTATACCCAG AAGTCACTCTCTCTGAGCCCTGGAAAGAAGGATCCCAAGTTTTGGG TCCTCGTGGTGGTCGGGGGTGTGCTGGCCTGCTACAGCTTGCTGG TCACAGTGGCCTTCATCATCTTCTGGGTGCGCTCCAAGAGGAGCC GGCTGCTTCACAGTGATTACATGAACATGACCCCCAGGAGGCCAG GACCCACCAGGAAGCACTACCAGCCCTACGCTCCCCCGCGGGAC TTTGCTGCTTACCGCAGCAGGGTCAAATTTTCTAGATCTGCAGATG CGCCGGCCTATCAAcaaggccagaaccagctcTATAACGAGCTCAATCTA GGACGAAGAGAGGAGTACGATGTTTTGGACAAGAGACGTGGCCG GGACCCTGAGATGGGGGGAAAGCCGAGAAGGAAGAACCCTCAGG AAGGCCTGTACAATGAACTGCAGAAAGATAAGATGGCGGAGGCCT ACAGTGAGATTGGGATGAAAGGCGAGCGCCGGAGGGGCAAGGGG CACGATGGCCTTTACCAGGGTCTCAGTACAGCCACCAAGGACACC TACGACGCCCTTCACATGCAGGCCCTGCCCCCTCGCTAAtaagaattct aactagagctcgctgatcagcctcgactgtgccttctagttgccagccatctgttgtttgcccctccccc gtgccttccttgaccctggaaggtgccactcccactgtcctttcctaataaaatgaggaaattgcatcg cattgtctgagtaggtgtcattctattctggggggtggggtggggcaggacagcaagggggaggat tgggaagagaatagcaggcatgctggggactatggacttcaagagcaacagtgctgtggcctgg agcaacaaatctgactttgcatgtgcaaacgccttcaacaacagcattattccagaagacaccttctt ccccagcccaggtaagggcagctttggtgccttcgcaggctgtttccttgcttcaggaatggccaggt tctgcccagagctctggtcaatgatgtctaaaactcctctgattggtggtctcggccttatccattgcca ccaaaaccctctttttactaagaaacagtgagccttgttctggcagtccagagaatgacacgggaa aaaagcagatgaagagaaggtggcaggagagggcacgtggcccagcctcagtctctccaactg agttcctgcctgcctgcctttgctcagact

[0214] TTTCAGGTTTCCTTGAGTGGCAGGCCAGGCCTGGCCGTGAACGTT CACTGAAATCATGGCCTCTTGGCCAAGATTGATAGCTTGTGCCTGT CCCTGAGTCCCAGTCCATCACGAGCAGCTGGTTTCTAAGATGCTAT HDRT3

[0215] 40 TTCCCGTATAAAGCATGAGACCGTGACTTGCCAGCCCCACAGAGC (T1 / T9) CCCGCCCTTGTCCATCACTGGCATCTGGACTCCAGCCTGGGTTGG GGCAAAGAGGGAAATGAGATCATGTCCTAACCCTGATCCTCTTGTC CCACAGATATCCAGAACCCTGACCCTGCCTCCGGATCCGGAGCCA CCAACTTCAGCCTGCTGAAGCAGGCCGGCGACGTGGAAGAAAATC CTGGGCCCATGGCTCTTCCTGTGACTGCCCTTCTGCTGCCCCTGG CTCTCCTGCTCCATGCTGCCCGGCCAGACATCCAGATGACACAGA CAACCAGCAGCCTCAGCGCCAGCCTGGGGGACAGAGTGACCATTA GCTGCCGGGCCTCTCAGGACATCAGCAAATACCTGAACTGGTACC AGCAGAAACCAGATGGCACTGTCAAGCTGCTGATTTACCACACCTC CAGGCTCCACAGCGGCGTGCCCAGTCGCTTCAGCGGCAGTGGGA GCGGGACAGATTATTCCCTCACAATCTCCAACCTGGAGCAGGAAG ATATTGCCACATACTTCTGCCAGCAAGGCAACACCCTGCCATACAC ATTTGGAGGCGGCACCAAATTGGAGATCACCGGCGGTGGTGGATC TGGAGGAGGGGGCAGCGGAGGTGGCGGCTCTGAGGTGAAACTGC AGGAGAGTGGCCCTGGCCTGGTGGCTCCCAGCCAGAGCCTTTCT GTCACCTGCACCGTGTCTGGGGTGTCCCTGCCTGACTATGGAGTC TCCTGGATCCGGCAGCCTCCAAGAAAAGGACTGGAATGGCTGGGC GTCATCTGGGGAAGTGAGACCACCTACTATAATTCAGCCCTCAAGT CCCGGCTCACCATCATTAAGGACAACTCCAAATCCCAGGTGTTCCT GAAGATGAATTCTCTCCAGACTGATGACACAGCCATCTACTACTGT

[0216] GCCAAGCATTATTATTACGGCGGGTCCTATGCCATGGACTACTGG GGCCAGGGGACCAGTGTCACTGTTTCTTCTGAAAGCAAGTATGGC CCCCCATGCCCTCCCTGCCCCGGACAGCCGCGGGAGCCCCAGGT CTACACCCTCCCACCCTCCAGAGATGAGCTGACCAAGAACCAGGT TTCACTGACATGCCTGGTGAAGGGCTTCTACCCCTCTGACATTGCT GTGGAGTGGGAGAGCAATGGGCAGCCAGAGAACAACTACAAGAC CACACCTCCTGTCCTGGACAGCGACGGCTCCTTCTTCCTCTACAGT AAGCTCACTGTGGACAAGAGCCGCTGGCAGCAGGGCAATGTCTTC TCCTGCAGCGTGATGCACGAGGCCCTGCACAATGCCTATACCCAG AAGTCACTCTCTCTGAGCCCTGGAAAGAAGGATCCCAAGTTTTGGG TCCTCGTGGTGGTCGGGGGTGTGCTGGCCTGCTACAGCTTGCTGG TCACAGTGGCCTTCATCATCTTCTGGGTGCGCTCCAAGAGGAGCC GGCTGCTTCACAGTGATTACATGAACATGACCCCCAGGAGGCCAG GACCCACCAGGAAGCACTACCAGCCCTACGCTCCCCCGCGGGAC TTTGCTGCTTACCGCAGCAGGGTCAAATTTTCTAGATCTGCAGATG CGCCGGCCTATCAAcaaggccagaaccagctcTATAACGAGCTCAATCTA GGACGAAGAGAGGAGTACGATGTTTTGGACAAGAGACGTGGCCG GGACCCTGAGATGGGGGGAAAGCCGAGAAGGAAGAACCCTCAGG AAGGCCTGTACAATGAACTGCAGAAAGATAAGATGGCGGAGGCCT ACAGTGAGATTGGGATGAAAGGCGAGCGCCGGAGGGGCAAGGGG

[0217] CACGATGGCCTTTACCAGGGTCTCAGTACAGCCACCAAGGACACC TACGACGCCCTTCACATGCAGGCCCTGCCCCCTCGCTAAtaagaattct aactagagctcgctgatcagcctcgactgtgccttctagttgccagccatctgttgtttgcccctccccc gtgccttccttgaccctggaaggtgccactcccactgtcctttcctaataaaatgaggaaattgcatcg cattgtctgagtaggtgtcattctattctggggggtggggtggggcaggacagcaagggggaggat tgggaagagaatagcaggcatgctggggagcctggagcaacaaatctgactttgcatgtgcaaa cgccttcaacaacagcattattccagaagacaccttcttccccagcccaggtaagggcagctttggt gccttcgcaggctgtttccttgcttcaggaatggccaggttctgcccagagctctggtcaatgatgtcta aaactcctctgattggtggtctcggccttatccattgccaccaaaaccctctttttactaagaaacagtg agccttgttctggcagtccagagaatgacacgggaaaaaagcagatgaagagaaggtggcagg agagggcacgtggcccagcctcagtctctccaactgagttcctgcctgcctgcctttgctcagactgtt tg cccctta ctg ctcttcta g gcctc

[0218] TTTCAGGTTTCCTTGAGTGGCAGGCCAGGCCTGGCCGTGAACGTT CACTGAAATCATGGCCTCTTGGCCAAGATTGATAGCTTGTGCCTGT CCCTGAGTCCCAGTCCATCACGAGCAGCTGGTTTCTAAGATGCTAT TTCCCGTATAAAGCATGAGACCGTGACTTGCCAGCCCCACAGAGC HDRT4

[0219] 41 CCCGCCCTTGTCCATCACTGGCATCTGGACTCCAGCCTGGGTTGG (T1 / T10) GGCAAAGAGGGAAATGAGATCATGTCCTAACCCTGATCCTCTTGTC CCACAGATATCCAGAACCCTGACCCTGCCTCCGGATCCGGAGCCA CCAACTTCAGCCTGCTGAAGCAGGCCGGCGACGTGGAAGAAAATC CTGGGCCCATGGCTCTTCCTGTGACTGCCCTTCTGCTGCCCCTGG CTCTCCTGCTCCATGCTGCCCGGCCAGACATCCAGATGACACAGA CAACCAGCAGCCTCAGCGCCAGCCTGGGGGACAGAGTGACCATTA GCTGCCGGGCCTCTCAGGACATCAGCAAATACCTGAACTGGTACC AGCAGAAACCAGATGGCACTGTCAAGCTGCTGATTTACCACACCTC CAGGCTCCACAGCGGCGTGCCCAGTCGCTTCAGCGGCAGTGGGA GCGGGACAGATTATTCCCTCACAATCTCCAACCTGGAGCAGGAAG ATATTGCCACATACTTCTGCCAGCAAGGCAACACCCTGCCATACAC ATTTGGAGGCGGCACCAAATTGGAGATCACCGGCGGTGGTGGATC TGGAGGAGGGGGCAGCGGAGGTGGCGGCTCTGAGGTGAAACTGC AGGAGAGTGGCCCTGGCCTGGTGGCTCCCAGCCAGAGCCTTTCT GTCACCTGCACCGTGTCTGGGGTGTCCCTGCCTGACTATGGAGTC TCCTGGATCCGGCAGCCTCCAAGAAAAGGACTGGAATGGCTGGGC GTCATCTGGGGAAGTGAGACCACCTACTATAATTCAGCCCTCAAGT CCCGGCTCACCATCATTAAGGACAACTCCAAATCCCAGGTGTTCCT GAAGATGAATTCTCTCCAGACTGATGACACAGCCATCTACTACTGT GCCAAGCATTATTATTACGGCGGGTCCTATGCCATGGACTACTGG GGCCAGGGGACCAGTGTCACTGTTTCTTCTGAAAGCAAGTATGGC CCCCCATGCCCTCCCTGCCCCGGACAGCCGCGGGAGCCCCAGGT CTACACCCTCCCACCCTCCAGAGATGAGCTGACCAAGAACCAGGT TTCACTGACATGCCTGGTGAAGGGCTTCTACCCCTCTGACATTGCT GTGGAGTGGGAGAGCAATGGGCAGCCAGAGAACAACTACAAGAC CACACCTCCTGTCCTGGACAGCGACGGCTCCTTCTTCCTCTACAGT AAGCTCACTGTGGACAAGAGCCGCTGGCAGCAGGGCAATGTCTTC TCCTGCAGCGTGATGCACGAGGCCCTGCACAATGCCTATACCCAG AAGTCACTCTCTCTGAGCCCTGGAAAGAAGGATCCCAAGTTTTGGG TCCTCGTGGTGGTCGGGGGTGTGCTGGCCTGCTACAGCTTGCTGG TCACAGTGGCCTTCATCATCTTCTGGGTGCGCTCCAAGAGGAGCC GGCTGCTTCACAGTGATTACATGAACATGACCCCCAGGAGGCCAG GACCCACCAGGAAGCACTACCAGCCCTACGCTCCCCCGCGGGAC TTTGCTGCTTACCGCAGCAGGGTCAAATTTTCTAGATCTGCAGATG CGCCGGCCTATCAAcaaggccagaaccagctcTATAACGAGCTCAATCTA GGACGAAGAGAGGAGTACGATGTTTTGGACAAGAGACGTGGCCG GGACCCTGAGATGGGGGGAAAGCCGAGAAGGAAGAACCCTCAGG AAGGCCTGTACAATGAACTGCAGAAAGATAAGATGGCGGAGGCCT ACAGTGAGATTGGGATGAAAGGCGAGCGCCGGAGGGGCAAGGGG CACGATGGCCTTTACCAGGGTCTCAGTACAGCCACCAAGGACACC TACGACGCCCTTCACATGCAGGCCCTGCCCCCTCGCTAAtaagaattct aactagagctcgctgatcagcctcgactgtgccttctagttgccagccatctgttgtttgcccctccccc gtgccttccttgaccctggaaggtgccactcccactgtcctttcctaataaaatgaggaaattgcatcg cattgtctgagtaggtgtcattctattctggggggtggggtggggcaggacagcaagggggaggat tgggaagagaatagcaggcatgctggggaaaggattctgatgtgtatatcacagacaaaactgtg ctagacatgaggtctatggacttcaagagcaacagtgctgtggcctggagcaacaaatctgactttg catgtgcaaacgccttcaacaacagcattattccagaagacaccttcttccccagcccaggtaagg gcagctttggtgccttcgcaggctgtttccttgcttcaggaatggccaggttctgcccagagctctggtc aatgatgtctaaaactcctctgattggtggtctcggccttatccattgccaccaaaaccctctttttacta agaaacagtgagccttgttctggcagtccagagaatgacacgggaaaaaagcagatg

[0220] TTTCAGGTTTCCTTGAGTGGCAGGCCAGGCCTGGCCGTGAACGTT CACTGAAATCATGGCCTCTTGGCCAAGATTGATAGCTTGTGCCTGT CCCTGAGTCCCAGTCCATCACGAGCAGCTGGTTTCTAAGATGCTAT TTCCCGTATAAAGCATGAGACCGTGACTTGCCAGCCCCACAGAGC CCCGCCCTTGTCCATCACTGGCATCTGGACTCCAGCCTGGGTTGG

[0221] HDRT1

[0222] 42 GGCAAAGAGGGAAATGAGATCATGTCCTAACCCTGATCCTCTTGTC (T1 / T11) CCACAGATATCCAGAACCCTGACCCTGCCTCCGGATCCGGAGCCA CCAACTTCAGCCTGCTGAAGCAGGCCGGCGACGTGGAAGAAAATC CTGGGCCCATGGCTCTTCCTGTGACTGCCCTTCTGCTGCCCCTGG CTCTCCTGCTCCATGCTGCCCGGCCAGACATCCAGATGACACAGA CAACCAGCAGCCTCAGCGCCAGCCTGGGGGACAGAGTGACCATTA GCTGCCGGGCCTCTCAGGACATCAGCAAATACCTGAACTGGTACC AGCAGAAACCAGATGGCACTGTCAAGCTGCTGATTTACCACACCTC CAGGCTCCACAGCGGCGTGCCCAGTCGCTTCAGCGGCAGTGGGA GCGGGACAGATTATTCCCTCACAATCTCCAACCTGGAGCAGGAAG ATATTGCCACATACTTCTGCCAGCAAGGCAACACCCTGCCATACAC ATTTGGAGGCGGCACCAAATTGGAGATCACCGGCGGTGGTGGATC TGGAGGAGGGGGCAGCGGAGGTGGCGGCTCTGAGGTGAAACTGC AGGAGAGTGGCCCTGGCCTGGTGGCTCCCAGCCAGAGCCTTTCT GTCACCTGCACCGTGTCTGGGGTGTCCCTGCCTGACTATGGAGTC TCCTGGATCCGGCAGCCTCCAAGAAAAGGACTGGAATGGCTGGGC GTCATCTGGGGAAGTGAGACCACCTACTATAATTCAGCCCTCAAGT CCCGGCTCACCATCATTAAGGACAACTCCAAATCCCAGGTGTTCCT GAAGATGAATTCTCTCCAGACTGATGACACAGCCATCTACTACTGT GCCAAGCATTATTATTACGGCGGGTCCTATGCCATGGACTACTGG GGCCAGGGGACCAGTGTCACTGTTTCTTCTGAAAGCAAGTATGGC CCCCCATGCCCTCCCTGCCCCGGACAGCCGCGGGAGCCCCAGGT CTACACCCTCCCACCCTCCAGAGATGAGCTGACCAAGAACCAGGT TTCACTGACATGCCTGGTGAAGGGCTTCTACCCCTCTGACATTGCT GTGGAGTGGGAGAGCAATGGGCAGCCAGAGAACAACTACAAGAC CACACCTCCTGTCCTGGACAGCGACGGCTCCTTCTTCCTCTACAGT AAGCTCACTGTGGACAAGAGCCGCTGGCAGCAGGGCAATGTCTTC TCCTGCAGCGTGATGCACGAGGCCCTGCACAATGCCTATACCCAG AAGTCACTCTCTCTGAGCCCTGGAAAGAAGGATCCCAAGTTTTGGG TCCTCGTGGTGGTCGGGGGTGTGCTGGCCTGCTACAGCTTGCTGG TCACAGTGGCCTTCATCATCTTCTGGGTGCGCTCCAAGAGGAGCC GGCTGCTTCACAGTGATTACATGAACATGACCCCCAGGAGGCCAG GACCCACCAGGAAGCACTACCAGCCCTACGCTCCCCCGCGGGAC TTTGCTGCTTACCGCAGCAGGGTCAAATTTTCTAGATCTGCAGATG CGCCGGCCTATCAAcaaggccagaaccagctcTATAACGAGCTCAATCTA GGACGAAGAGAGGAGTACGATGTTTTGGACAAGAGACGTGGCCG GGACCCTGAGATGGGGGGAAAGCCGAGAAGGAAGAACCCTCAGG AAGGCCTGTACAATGAACTGCAGAAAGATAAGATGGCGGAGGCCT ACAGTGAGATTGGGATGAAAGGCGAGCGCCGGAGGGGCAAGGGG CACGATGGCCTTTACCAGGGTCTCAGTACAGCCACCAAGGACACC TACGACGCCCTTCACATGCAGGCCCTGCCCCCTCGCTAAtaagaattct aactagagctcgctgatcagcctcgactgtgccttctagttgccagccatctgttgtttgcccctccccc gtgccttccttgaccctggaaggtgccactcccactgtcctttcctaataaaatgaggaaattgcatcg cattgtctgagtaggtgtcattctattctggggggtggggtggggcaggacagcaagggggaggat tgggaagagaatagcaggcatgctggggactttggtgccttcgcaggctgtttccttgcttcaggaat ggccaggttctgcccagagctctggtcaatgatgtctaaaactcctctgattggtggtctcggccttatc cattgccaccaaaaccctctttttactaagaaacagtgagccttgttctggcagtccagagaatgac acgggaaaaaagcagatgaagagaaggtggcaggagagggcacgtggcccagcctcagtct ctccaactgagttcctgcctgcctgcctttgctcagactgtttgccccttactgctcttctaggcctcattct aagccccttctccaagttgcctctccttatttctccctgtctgccaaaaaatctttcccagctcactaagt cagtctcacgcagtcactc TTTCAGGTTTCCTTGAGTGGCAGGCCAGGCCTGGCCGTGAACGTT CACTGAAATCATGGCCTCTTGGCCAAGATTGATAGCTTGTGCCTGT CCCTGAGTCCCAGTCCATCACGAGCAGCTGGTTTCTAAGATGCTAT TTCCCGTATAAAGCATGAGACCGTGACTTGCCAGCCCCACAGAGC CCCGCCCTTGTCCATCACTGGCATCTGGACTCCAGCCTGGGTTGG GGCAAAGAGGGAAATGAGATCATGTCCTAACCCTGATCCTCTTGTC

[0223] HDRT1

[0224] 43 CCACAGATATCCAGAACCCTGACCCTGCCTCCGGATCCGGAGCCA (T1 / T12) CCAACTTCAGCCTGCTGAAGCAGGCCGGCGACGTGGAAGAAAATC CTGGGCCCATGGCTCTTCCTGTGACTGCCCTTCTGCTGCCCCTGG CTCTCCTGCTCCATGCTGCCCGGCCAGACATCCAGATGACACAGA CAACCAGCAGCCTCAGCGCCAGCCTGGGGGACAGAGTGACCATTA GCTGCCGGGCCTCTCAGGACATCAGCAAATACCTGAACTGGTACC AGCAGAAACCAGATGGCACTGTCAAGCTGCTGATTTACCACACCTC CAGGCTCCACAGCGGCGTGCCCAGTCGCTTCAGCGGCAGTGGGA GCGGGACAGATTATTCCCTCACAATCTCCAACCTGGAGCAGGAAG ATATTGCCACATACTTCTGCCAGCAAGGCAACACCCTGCCATACAC ATTTGGAGGCGGCACCAAATTGGAGATCACCGGCGGTGGTGGATC TGGAGGAGGGGGCAGCGGAGGTGGCGGCTCTGAGGTGAAACTGC AGGAGAGTGGCCCTGGCCTGGTGGCTCCCAGCCAGAGCCTTTCT GTCACCTGCACCGTGTCTGGGGTGTCCCTGCCTGACTATGGAGTC TCCTGGATCCGGCAGCCTCCAAGAAAAGGACTGGAATGGCTGGGC GTCATCTGGGGAAGTGAGACCACCTACTATAATTCAGCCCTCAAGT CCCGGCTCACCATCATTAAGGACAACTCCAAATCCCAGGTGTTCCT GAAGATGAATTCTCTCCAGACTGATGACACAGCCATCTACTACTGT GCCAAGCATTATTATTACGGCGGGTCCTATGCCATGGACTACTGG GGCCAGGGGACCAGTGTCACTGTTTCTTCTGAAAGCAAGTATGGC CCCCCATGCCCTCCCTGCCCCGGACAGCCGCGGGAGCCCCAGGT CTACACCCTCCCACCCTCCAGAGATGAGCTGACCAAGAACCAGGT TTCACTGACATGCCTGGTGAAGGGCTTCTACCCCTCTGACATTGCT GTGGAGTGGGAGAGCAATGGGCAGCCAGAGAACAACTACAAGAC CACACCTCCTGTCCTGGACAGCGACGGCTCCTTCTTCCTCTACAGT AAGCTCACTGTGGACAAGAGCCGCTGGCAGCAGGGCAATGTCTTC TCCTGCAGCGTGATGCACGAGGCCCTGCACAATGCCTATACCCAG AAGTCACTCTCTCTGAGCCCTGGAAAGAAGGATCCCAAGTTTTGGG TCCTCGTGGTGGTCGGGGGTGTGCTGGCCTGCTACAGCTTGCTGG TCACAGTGGCCTTCATCATCTTCTGGGTGCGCTCCAAGAGGAGCC GGCTGCTTCACAGTGATTACATGAACATGACCCCCAGGAGGCCAG GACCCACCAGGAAGCACTACCAGCCCTACGCTCCCCCGCGGGAC TTTGCTGCTTACCGCAGCAGGGTCAAATTTTCTAGATCTGCAGATG CGCCGGCCTATCAAcaaggccagaaccagctcTATAACGAGCTCAATCTA GGACGAAGAGAGGAGTACGATGTTTTGGACAAGAGACGTGGCCG GGACCCTGAGATGGGGGGAAAGCCGAGAAGGAAGAACCCTCAGG AAGGCCTGTACAATGAACTGCAGAAAGATAAGATGGCGGAGGCCT ACAGTGAGATTGGGATGAAAGGCGAGCGCCGGAGGGGCAAGGGG CACGATGGCCTTTACCAGGGTCTCAGTACAGCCACCAAGGACACC TACGACGCCCTTCACATGCAGGCCCTGCCCCCTCGCTAAtaagaattct aactagagctcgctgatcagcctcgactgtgccttctagttgccagccatctgttgtttgcccctccccc gtgccttccttgaccctggaaggtgccactcccactgtcctttcctaataaaatgaggaaattgcatcg cattgtctgagtaggtgtcattctattctggggggtggggtggggcaggacagcaagggggaggat tgggaagagaatagcaggcatgctggggactttggtgccttcgcaggctgtttccttgcttcaggaat ggccaggttctgcccagagctctggtcaatgatgtctaaaactcctctgattggtggtctcggccttatc cattgccaccaaaaccctctttttactaagaaacagtgagccttgttctggcagtccagagaatgac acgggaaaaaagcagatgaagagaaggtggcaggagagggcacgtggcccagcctcagtct ctccaactgagttcctgcctgcctgcctttgctcagactgtttgccccttactgctcttctaggcctcattct aagccccttctccaagttgcctctccttatttctccctgtctgccaaaaaatctttcccagctcactaagt cagtctcacgcagtcactc _ cctattaaataaaagaataagcagtattattaagtagccctgcatttcaggtttccttgagtggcaggc caggcctggccgtgaacgttcactgaaatcatggcctcttggccaagattgatagcttgtgcctgtcc ctgagtcccagtccatcacgagcagctggtttctaagatgctatttcccgtataaagcatgagaccgt gacttgccagccccacagagccccgcccttgtccatcactggcatctggactccagcctgggttgg ggcaaagagggaaatgagatcatgtcctaaccctgatcctcttgtcccacagatatccagaaccct gaccctgccgtgtaccagctgagagactctaaatccagtgacaagtctgtctgcctattcaccGCC ACCAACTTCTCTCTTTTGAAGCAGGCCGGAGATGTGGAAGAAAATC

[0225] HDRT6

[0226] 44 CTGGGCCCATGGCTCTTCCTGTGACTGCCCTTCTGCTGCCCCTGG (T2 / T8) CTCTCCTGCTCCATGCTGCCCGGCCAGACATCCAGATGACACAGA

[0227] CAACCAGCAGCCTCAGCGCCAGCCTGGGGGACAGAGTGACCATTA GCTGCCGGGCCTCTCAGGACATCAGCAAATACCTGAACTGGTACC AGCAGAAACCAGATGGCACTGTCAAGCTGCTGATTTACCACACCTC CAGGCTCCACAGCGGCGTGCCCAGTCGCTTCAGCGGCAGTGGGA GCGGGACAGATTATTCCCTCACAATCTCCAACCTGGAGCAGGAAG ATATTGCCACATACTTCTGCCAGCAAGGCAACACCCTGCCATACAC ATTTGGAGGCGGCACCAAATTGGAGATCACCGGCGGTGGTGGATC TGGAGGAGGGGGCAGCGGAGGTGGCGGCTCTGAGGTGAAACTGC AGGAGAGTGGCCCTGGCCTGGTGGCTCCCAGCCAGAGCCTTTCT GTCACCTGCACCGTGTCTGGGGTGTCCCTGCCTGACTATGGAGTC TCCTGGATCCGGCAGCCTCCAAGAAAAGGACTGGAATGGCTGGGC GTCATCTGGGGAAGTGAGACCACCTACTATAATTCAGCCCTCAAGT CCCGGCTCACCATCATTAAGGACAACTCCAAATCCCAGGTGTTCCT GAAGATGAATTCTCTCCAGACTGATGACACAGCCATCTACTACTGT GCCAAGCATTATTATTACGGCGGGTCCTATGCCATGGACTACTGG GGCCAGGGGACCAGTGTCACTGTTTCTTCTGAAAGCAAGTATGGC CCCCCATGCCCTCCCTGCCCCGGACAGCCGCGGGAGCCCCAGGT CTACACCCTCCCACCCTCCAGAGATGAGCTGACCAAGAACCAGGT TTCACTGACATGCCTGGTGAAGGGCTTCTACCCCTCTGACATTGCT GTGGAGTGGGAGAGCAATGGGCAGCCAGAGAACAACTACAAGAC CACACCTCCTGTCCTGGACAGCGACGGCTCCTTCTTCCTCTACAGT AAGCTCACTGTGGACAAGAGCCGCTGGCAGCAGGGCAATGTCTTC TCCTGCAGCGTGATGCACGAGGCCCTGCACAATGCCTATACCCAG AAGTCACTCTCTCTGAGCCCTGGAAAGAAGGATCCCAAGTTTTGGG TCCTCGTGGTGGTCGGGGGTGTGCTGGCCTGCTACAGCTTGCTGG TCACAGTGGCCTTCATCATCTTCTGGGTGCGCTCCAAGAGGAGCC GGCTGCTTCACAGTGATTACATGAACATGACCCCCAGGAGGCCAG GACCCACCAGGAAGCACTACCAGCCCTACGCTCCCCCGCGGGAC TTTGCTGCTTACCGCAGCAGGGTCAAATTTTCTAGATCTGCAGATG CGCCGGCCTATCAAcaaggccagaaccagctcTATAACGAGCTCAATCTA GGACGAAGAGAGGAGTACGATGTTTTGGACAAGAGACGTGGCCG GGACCCTGAGATGGGGGGAAAGCCGAGAAGGAAGAACCCTCAGG AAGGCCTGTACAATGAACTGCAGAAAGATAAGATGGCGGAGGCCT ACAGTGAGATTGGGATGAAAGGCGAGCGCCGGAGGGGCAAGGGG CACGATGGCCTTTACCAGGGTCTCAGTACAGCCACCAAGGACACC TACGACGCCCTTCACATGCAGGCCCTGCCCCCTCGCTAAtaagaattct aactagagctcgctgatcagcctcgactgtgccttctagttgccagccatctgttgtttgcccctccccc gtgccttccttgaccctggaaggtgccactcccactgtcctttcctaataaaatgaggaaattgcatcg cattgtctgagtaggtgtcattctattctggggggtggggtggggcaggacagcaagggggaggat tgggaagagaatagcaggcatgctggggactatggacttcaagagcaacagtgctgtggcctgg agcaacaaatctgactttgcatgtgcaaacgccttcaacaacagcattattccagaagacaccttctt ccccagcccaggtaagggcagctttggtgccttcgcaggctgtttccttgcttcaggaatggccaggt tctgcccagagctctggtcaatgatgtctaaaactcctctgattggtggtctcggccttatccattgcca ccaaaaccctctttttactaagaaacagtgagccttgttctggcagtccagagaatgacacgggaa aaaagcagatgaagagaaggtggcaggagagggcacgtggcccagcctcagtctctccaactg agttcctgcctgcctgcctttgctcagact _ cctattaaataaaagaataagcagtattattaagtagccctgcatttcaggtttccttgagtggcaggc caggcctggccgtgaacgttcactgaaatcatggcctcttggccaagattgatagcttgtgcctgtcc ctgagtcccagtccatcacgagcagctggtttctaagatgctatttcccgtataaagcatgagaccgt gacttgccagccccacagagccccgcccttgtccatcactggcatctggactccagcctgggttgg ggcaaagagggaaatgagatcatgtcctaaccctgatcctcttgtcccacagatatccagaaccct gaccctgccgtgtaccagctgagagactctaaatccagtgacaagtctgtctgcctattcaccGCC ACCAACTTCTCTCTTTTGAAGCAGGCCGGAGATGTGGAAGAAAATC CTGGGCCCATGGCTCTTCCTGTGACTGCCCTTCTGCTGCCCCTGG CTCTCCTGCTCCATGCTGCCCGGCCAGACATCCAGATGACACAGA HDRT7

[0228] 45 CAACCAGCAGCCTCAGCGCCAGCCTGGGGGACAGAGTGACCATTA (T2 / T9) GCTGCCGGGCCTCTCAGGACATCAGCAAATACCTGAACTGGTACC AGCAGAAACCAGATGGCACTGTCAAGCTGCTGATTTACCACACCTC CAGGCTCCACAGCGGCGTGCCCAGTCGCTTCAGCGGCAGTGGGA GCGGGACAGATTATTCCCTCACAATCTCCAACCTGGAGCAGGAAG ATATTGCCACATACTTCTGCCAGCAAGGCAACACCCTGCCATACAC ATTTGGAGGCGGCACCAAATTGGAGATCACCGGCGGTGGTGGATC TGGAGGAGGGGGCAGCGGAGGTGGCGGCTCTGAGGTGAAACTGC AGGAGAGTGGCCCTGGCCTGGTGGCTCCCAGCCAGAGCCTTTCT GTCACCTGCACCGTGTCTGGGGTGTCCCTGCCTGACTATGGAGTC TCCTGGATCCGGCAGCCTCCAAGAAAAGGACTGGAATGGCTGGGC GTCATCTGGGGAAGTGAGACCACCTACTATAATTCAGCCCTCAAGT CCCGGCTCACCATCATTAAGGACAACTCCAAATCCCAGGTGTTCCT GAAGATGAATTCTCTCCAGACTGATGACACAGCCATCTACTACTGT GCCAAGCATTATTATTACGGCGGGTCCTATGCCATGGACTACTGG GGCCAGGGGACCAGTGTCACTGTTTCTTCTGAAAGCAAGTATGGC CCCCCATGCCCTCCCTGCCCCGGACAGCCGCGGGAGCCCCAGGT CTACACCCTCCCACCCTCCAGAGATGAGCTGACCAAGAACCAGGT TTCACTGACATGCCTGGTGAAGGGCTTCTACCCCTCTGACATTGCT GTGGAGTGGGAGAGCAATGGGCAGCCAGAGAACAACTACAAGAC CACACCTCCTGTCCTGGACAGCGACGGCTCCTTCTTCCTCTACAGT AAGCTCACTGTGGACAAGAGCCGCTGGCAGCAGGGCAATGTCTTC TCCTGCAGCGTGATGCACGAGGCCCTGCACAATGCCTATACCCAG AAGTCACTCTCTCTGAGCCCTGGAAAGAAGGATCCCAAGTTTTGGG TCCTCGTGGTGGTCGGGGGTGTGCTGGCCTGCTACAGCTTGCTGG TCACAGTGGCCTTCATCATCTTCTGGGTGCGCTCCAAGAGGAGCC GGCTGCTTCACAGTGATTACATGAACATGACCCCCAGGAGGCCAG GACCCACCAGGAAGCACTACCAGCCCTACGCTCCCCCGCGGGAC TTTGCTGCTTACCGCAGCAGGGTCAAATTTTCTAGATCTGCAGATG CGCCGGCCTATCAAcaaggccagaaccagctcTATAACGAGCTCAATCTA GGACGAAGAGAGGAGTACGATGTTTTGGACAAGAGACGTGGCCG GGACCCTGAGATGGGGGGAAAGCCGAGAAGGAAGAACCCTCAGG AAGGCCTGTACAATGAACTGCAGAAAGATAAGATGGCGGAGGCCT ACAGTGAGATTGGGATGAAAGGCGAGCGCCGGAGGGGCAAGGGG CACGATGGCCTTTACCAGGGTCTCAGTACAGCCACCAAGGACACC TACGACGCCCTTCACATGCAGGCCCTGCCCCCTCGCTAAtaagaattct aactagagctcgctgatcagcctcgactgtgccttctagttgccagccatctgttgtttgcccctccccc gtgccttccttgaccctggaaggtgccactcccactgtcctttcctaataaaatgaggaaattgcatcg cattgtctgagtaggtgtcattctattctggggggtggggtggggcaggacagcaagggggaggat tgggaagagaatagcaggcatgctggggagcctggagcaacaaatctgactttgcatgtgcaaa cgccttcaacaacagcattattccagaagacaccttcttccccagcccaggtaagggcagctttggt gccttcgcaggctgtttccttgcttcaggaatggccaggttctgcccagagctctggtcaatgatgtcta aaactcctctgattggtggtctcggccttatccattgccaccaaaaccctctttttactaagaaacagtg agccttgttctggcagtccagagaatgacacgggaaaaaagcagatgaagagaaggtggcagg agagggcacgtggcccagcctcagtctctccaactgagttcctgcctgcctgcctttgctcagactgtt tgccccttactgctcttctaggcctc _ cctattaaataaaagaataagcagtattattaagtagccctgcatttcaggtttccttgagtggcaggc caggcctggccgtgaacgttcactgaaatcatggcctcttggccaagattgatagcttgtgcctgtcc ctgagtcccagtccatcacgagcagctggtttctaagatgctatttcccgtataaagcatgagaccgt gacttgccagccccacagagccccgcccttgtccatcactggcatctggactccagcctgggttgg ggcaaagagggaaatgagatcatgtcctaaccctgatcctcttgtcccacagatatccagaaccct gaccctgccgtgtaccagctgagagactctaaatccagtgacaagtctgtctgcctattcaccGCC ACCAACTTCTCTCTTTTGAAGCAGGCCGGAGATGTGGAAGAAAATC

[0229] CTGGGCCCATGGCTCTTCCTGTGACTGCCCTTCTGCTGCCCCTGG CTCTCCTGCTCCATGCTGCCCGGCCAGACATCCAGATGACACAGA CAACCAGCAGCCTCAGCGCCAGCCTGGGGGACAGAGTGACCATTA

[0230] HDRT8

[0231] 46 GCTGCCGGGCCTCTCAGGACATCAGCAAATACCTGAACTGGTACC (T2 / T10) AGCAGAAACCAGATGGCACTGTCAAGCTGCTGATTTACCACACCTC CAGGCTCCACAGCGGCGTGCCCAGTCGCTTCAGCGGCAGTGGGA GCGGGACAGATTATTCCCTCACAATCTCCAACCTGGAGCAGGAAG ATATTGCCACATACTTCTGCCAGCAAGGCAACACCCTGCCATACAC ATTTGGAGGCGGCACCAAATTGGAGATCACCGGCGGTGGTGGATC TGGAGGAGGGGGCAGCGGAGGTGGCGGCTCTGAGGTGAAACTGC AGGAGAGTGGCCCTGGCCTGGTGGCTCCCAGCCAGAGCCTTTCT GTCACCTGCACCGTGTCTGGGGTGTCCCTGCCTGACTATGGAGTC TCCTGGATCCGGCAGCCTCCAAGAAAAGGACTGGAATGGCTGGGC GTCATCTGGGGAAGTGAGACCACCTACTATAATTCAGCCCTCAAGT CCCGGCTCACCATCATTAAGGACAACTCCAAATCCCAGGTGTTCCT GAAGATGAATTCTCTCCAGACTGATGACACAGCCATCTACTACTGT GCCAAGCATTATTATTACGGCGGGTCCTATGCCATGGACTACTGG GGCCAGGGGACCAGTGTCACTGTTTCTTCTGAAAGCAAGTATGGC CCCCCATGCCCTCCCTGCCCCGGACAGCCGCGGGAGCCCCAGGT CTACACCCTCCCACCCTCCAGAGATGAGCTGACCAAGAACCAGGT TTCACTGACATGCCTGGTGAAGGGCTTCTACCCCTCTGACATTGCT GTGGAGTGGGAGAGCAATGGGCAGCCAGAGAACAACTACAAGAC CACACCTCCTGTCCTGGACAGCGACGGCTCCTTCTTCCTCTACAGT AAGCTCACTGTGGACAAGAGCCGCTGGCAGCAGGGCAATGTCTTC TCCTGCAGCGTGATGCACGAGGCCCTGCACAATGCCTATACCCAG AAGTCACTCTCTCTGAGCCCTGGAAAGAAGGATCCCAAGTTTTGGG TCCTCGTGGTGGTCGGGGGTGTGCTGGCCTGCTACAGCTTGCTGG TCACAGTGGCCTTCATCATCTTCTGGGTGCGCTCCAAGAGGAGCC GGCTGCTTCACAGTGATTACATGAACATGACCCCCAGGAGGCCAG GACCCACCAGGAAGCACTACCAGCCCTACGCTCCCCCGCGGGAC TTTGCTGCTTACCGCAGCAGGGTCAAATTTTCTAGATCTGCAGATG CGCCGGCCTATCAAcaaggccagaaccagctcTATAACGAGCTCAATCTA GGACGAAGAGAGGAGTACGATGTTTTGGACAAGAGACGTGGCCG GGACCCTGAGATGGGGGGAAAGCCGAGAAGGAAGAACCCTCAGG AAGGCCTGTACAATGAACTGCAGAAAGATAAGATGGCGGAGGCCT ACAGTGAGATTGGGATGAAAGGCGAGCGCCGGAGGGGCAAGGGG CACGATGGCCTTTACCAGGGTCTCAGTACAGCCACCAAGGACACC TACGACGCCCTTCACATGCAGGCCCTGCCCCCTCGCTAAtaagaattct aactagagctcgctgatcagcctcgactgtgccttctagttgccagccatctgttgtttgcccctccccc gtgccttccttgaccctggaaggtgccactcccactgtcctttcctaataaaatgaggaaattgcatcg cattgtctgagtaggtgtcattctattctggggggtggggtggggcaggacagcaagggggaggat tgggaagagaatagcaggcatgctggggaaaggattctgatgtgtatatcacagacaaaactgtg ctagacatgaggtctatggacttcaagagcaacagtgctgtggcctggagcaacaaatctgactttg catgtgcaaacgccttcaacaacagcattattccagaagacaccttcttccccagcccaggtaagg gcagctttggtgccttcgcaggctgtttccttgcttcaggaatggccaggttctgcccagagctctggtc aatgatgtctaaaactcctctgattggtggtctcggccttatccattgccaccaaaaccctctttttacta agaaacagtgagccttgttctggcagtccagagaatgacacgggaaaaaagcagatg _ tgatagcttgtgcctgtccctgagtcccagtccatcacgagcagctggtttctaagatgctatttcccgt ataaagcatgagaccgtgacttgccagccccacagagccccgcccttgtccatcactggcatctgg actccagcctgggttggggcaaagagggaaatgagatcatgtcctaaccctgatcctcttgtcccac agatatccagaaccctgaccctgccgtgtaccagctgagagactctaaatccagtgacaagtctgt ctgcctattcaccgattttgattctcaaacaaatgtgtcacaaagtaaggattctgatgtgtatatcaca gacaaaactgtgctagacatgaggtctatggacttcaagagcaacagtgctgtggcctggagcG GATCCGGAGCCACCAACTTCAGCCTGCTGAAGCAGGCCGGCGAC GTGGAAGAAAATCCTGGGCCCATGGCTCTTCCTGTGACTGCCCTT CTGCTGCCCCTGGCTCTCCTGCTCCATGCTGCCCGGCCAGACATC CAGATGACACAGACAACCAGCAGCCTCAGCGCCAGCCTGGGGGA CAGAGTGACCATTAGCTGCCGGGCCTCTCAGGACATCAGCAAATA CCTGAACTGGTACCAGCAGAAACCAGATGGCACTGTCAAGCTGCT

[0232] HDRT9

[0233] 47 GATTTACCACACCTCCAGGCTCCACAGCGGCGTGCCCAGTCGCTT (T3 / T12) CAGCGGCAGTGGGAGCGGGACAGATTATTCCCTCACAATCTCCAA CCTGGAGCAGGAAGATATTGCCACATACTTCTGCCAGCAAGGCAA CACCCTGCCATACACATTTGGAGGCGGCACCAAATTGGAGATCAC CGGCGGTGGTGGATCTGGAGGAGGGGGCAGCGGAGGTGGCGGC TCTGAGGTGAAACTGCAGGAGAGTGGCCCTGGCCTGGTGGCTCC CAGCCAGAGCCTTTCTGTCACCTGCACCGTGTCTGGGGTGTCCCT GCCTGACTATGGAGTCTCCTGGATCCGGCAGCCTCCAAGAAAAGG ACTGGAATGGCTGGGCGTCATCTGGGGAAGTGAGACCACCTACTA TAATTCAGCCCTCAAGTCCCGGCTCACCATCATTAAGGACAACTCC AAATCCCAGGTGTTCCTGAAGATGAATTCTCTCCAGACTGATGACA CAGCCATCTACTACTGTGCCAAGCATTATTATTACGGCGGGTCCTA TGCCATGGACTACTGGGGCCAGGGGACCAGTGTCACTGTTTCTTC TGAAAGCAAGTATGGCCCCCCATGCCCTCCCTGCCCCGGACAGCC GCGGGAGCCCCAGGTCTACACCCTCCCACCCTCCAGAGATGAGCT GACCAAGAACCAGGTTTCACTGACATGCCTGGTGAAGGGCTTCTA CCCCTCTGACATTGCTGTGGAGTGGGAGAGCAATGGGCAGCCAGA GAACAACTACAAGACCACACCTCCTGTCCTGGACAGCGACGGCTC CTTCTTCCTCTACAGTAAGCTCACTGTGGACAAGAGCCGCTGGCA GCAGGGCAATGTCTTCTCCTGCAGCGTGATGCACGAGGCCCTGCA CAATGCCTATACCCAGAAGTCACTCTCTCTGAGCCCTGGAAAGAAG GATCCCAAGTTTTGGGTCCTCGTGGTGGTCGGGGGTGTGCTGGCC TGCTACAGCTTGCTGGTCACAGTGGCCTTCATCATCTTCTGGGTGC GCTCCAAGAGGAGCCGGCTGCTTCACAGTGATTACATGAACATGA CCCCCAGGAGGCCAGGACCCACCAGGAAGCACTACCAGCCCTAC GCTCCCCCGCGGGACTTTGCTGCTTACCGCAGCAGGGTCAAATTT T CT AGAT CT GCAGAT GCGCCGGCCTAT CAAcaaggccagaaccagctcTA TAACGAGCTCAATCTAGGACGAAGAGAGGAGTACGATGTTTTGGA CAAGAGACGTGGCCGGGACCCTGAGATGGGGGGAAAGCCGAGAA GGAAGAACCCTCAGGAAGGCCTGTACAATGAACTGCAGAAAGATA AGATGGCGGAGGCCTACAGTGAGATTGGGATGAAAGGCGAGCGC CGGAGGGGCAAGGGGCACGATGGCCTTTACCAGGGTCTCAGTAC AGCCACCAAGGACACCTACGACGCCCTTCACATGCAGGCCCTGCC CCCTCGCTAAtaagaattctaactagagctcgctgatcagcctcgactgtgccttctagttgcc agccatctgttgtttgcccctcccccgtgccttccttgaccctggaaggtgccactcccactgtcctttcc taataaaatgaggaaattgcatcgcattgtctgagtaggtgtcattctattctggggggtggggtggg gcaggacagcaagggggaggattgggaagagaatagcaggcatgctggggataagggcagc tttggtgccttcgcaggctgtttccttgcttcaggaatggccaggttctgcccagagctctggtcaatga tgtctaaaactcctctgattggtggtctcggccttatccattgccaccaaaaccctctttttactaagaaa cagtgagccttgttctggcagtccagagaatgacacgggaaaaaagcagatgaagagaaggtg gcaggagagggcacgtggcccagcctcagtctctccaactgagttcctgcctgcctgcctttgctca gactgtttgccccttactgctcttctaggcctcattctaagccccttctccaagttgcctctccttatttctcc ctgtctgccaaaaaatctttcccagctcactaagtcagtctcacgcagtcactc _ atctcatttgaccatcatttgacctattgtctccctgtgggtaggcctcagagccacacacctcaggcc aggagtaccattcatccagacgtgaacatcttcccgaggcttccagagttcttggttcacaccgggg ctaacatggctgggcttctgctgcagtggcaggagctctgtgcacagagaacagcctcatctgctcg ccttgtttccacctcccctcccattgccccaggttctttggccccacagcggccacatctgccgttggtg ccaataggttttccaggagctggttgaggtgggagggagggagagggttgtgatcaggctgaggc atggggattggatatagtctccgtgtcatgatttatttggtcagtcagtcctagtgcGaccctggggta atggggatgtgttctcgtcacGttgggcctggctgaccagctttatctcttggcacagaggcaggca gcggcGCTACCAATTTTTCTTTGTTGAAGCAGGCCGGGGATGTGGAG GAAAATCCGGGGCCAATGGCTCTTCCTGTGACTGCCCTTCTGCTG CCCCTGGCTCTCCTGCTCCATGCTGCCCGGCCAGACATCCAGATG ACACAGACAACCAGCAGCCTCAGCGCCAGCCTGGGGGACAGAGT GACCATTAGCTGCCGGGCCTCTCAGGACATCAGCAAATACCTGAA CTGGTACCAGCAGAAACCAGATGGCACTGTCAAGCTGCTGATTTAC CACACCTCCAGGCTCCACAGCGGCGTGCCCAGTCGCTTCAGCGG HDRT17

[0234] 48 CAGTGGGAGCGGGACAGATTATTCCCTCACAATCTCCAACCTGGA (Z2 / Z5) GCAGGAAGATATTGCCACATACTTCTGCCAGCAAGGCAACACCCT GCCATACACATTTGGAGGCGGCACCAAATTGGAGATCACCGGCGG TGGTGGATCTGGAGGAGGGGGCAGCGGAGGTGGCGGCTCTGAG GTGAAACTGCAGGAGAGTGGCCCTGGCCTGGTGGCTCCCAGCCA GAGCCTTTCTGTCACCTGCACCGTGTCTGGGGTGTCCCTGCCTGA CTATGGAGTCTCCTGGATCCGGCAGCCTCCAAGAAAAGGACTGGA ATGGCTGGGCGTCATCTGGGGAAGTGAGACCACCTACTATAATTC AGCCCTCAAGTCCCGGCTCACCATCATTAAGGACAACTCCAAATCC CAGGTGTTCCTGAAGATGAATTCTCTCCAGACTGATGACACAGCCA TCTACTACTGTGCCAAGCATTATTATTACGGCGGGTCCTATGCCAT GGACTACTGGGGCCAGGGGACCAGTGTCACTGTTTCTTCTGAAAG CAAGTATGGCCCCCCATGCCCTCCCTGCCCCGGACAGCCGCGGG AGCCCCAGGTCTACACCCTCCCACCCTCCAGAGATGAGCTGACCA AGAACCAGGTTTCACTGACATGCCTGGTGAAGGGCTTCTACCCCT CTGACATTGCTGTGGAGTGGGAGAGCAATGGGCAGCCAGAGAACA ACTACAAGACCACACCTCCTGTCCTGGACAGCGACGGCTCCTTCTT CCTCTACAGTAAGCTCACTGTGGACAAGAGCCGCTGGCAGCAGGG CAATGTCTTCTCCTGCAGCGTGATGCACGAGGCCCTGCACAATGC CTATACCCAGAAGTCACTCTCTCTGAGCCCTGGAAAGAAGGATCCC AAGTTTTGGGTCCTCGTGGTGGTCGGGGGTGTGCTGGCCTGCTAC AGCTTGCTGGTCACAGTGGCCTTCATCATCTTCTGGGTGCGCTCCA AGAGGAGCCGGCTGCTTCACAGTGATTACATGAACATGACCCCCA GGAGGCCAGGACCCACCAGGAAGCACTACCAGCCCTACGCTCCC CCGCGGGACTTTGCTGCTTACCGCAGCagagtgaaggtgggtaccactgggct ttgggaggagggcacggggtcccccacttgatggatgttcagaggggccttggtcttggaaggtctc aagctcgggtggtgcctggggcttggtatccaggagcaaagcaaggaccagccaagtgtgtgcct tgagtgggctgaggaggaggtggcagtgtctggctgagatggacagggtaggagggagagcct ggtgctaggcacctccatgacaagccgtacaaatgtgtgcacatcagagtgtcccagggaaggc gatgctactggtgacaaaggggcttacactcaggcagaggtccttctttccaagtgtgaatgaaggc catgttagcctttctcttgaaaaggccctttcctcatctgtaactggggaG _ atctcatttgaccatcatttgacctattgtctccctgtgggtaggcctcagagccacacacctcaggcc aggagtaccattcatccagacgtgaacatcttcccgaggcttccagagttcttggttcacaccgggg ctaacatggctgggcttctgctgcagtggcaggagctctgtgcacagagaacagcctcatctgctcg ccttgtttccacctcccctcccattgccccaggttctttggccccacagcggccacatctgccgttggtg ccaataggttttccaggagctggttgaggtgggagggagggagagggttgtgatcaggctgaggc atggggattggatatagtctccgtgtcatgatttatttggtcagtcagtcctagtgcGaccctggggta atggggatgtgttctcgtcacGttgggcctggctgaccagctttatctcttggcacagaggcaggca gcggcGCTACCAATTTTTCTTTGTTGAAGCAGGCCGGGGATGTGGAG GAAAATCCGGGGCCAATGGCTCTTCCTGTGACTGCCCTTCTGCTG CCCCTGGCTCTCCTGCTCCATGCTGCCCGGCCAGACATCCAGATG ACACAGACAACCAGCAGCCTCAGCGCCAGCCTGGGGGACAGAGT GACCATTAGCTGCCGGGCCTCTCAGGACATCAGCAAATACCTGAA CTGGTACCAGCAGAAACCAGATGGCACTGTCAAGCTGCTGATTTAC CACACCTCCAGGCTCCACAGCGGCGTGCCCAGTCGCTTCAGCGG CAGTGGGAGCGGGACAGATTATTCCCTCACAATCTCCAACCTGGA GCAGGAAGATATTGCCACATACTTCTGCCAGCAAGGCAACACCCT GCCATACACATTTGGAGGCGGCACCAAATTGGAGATCACCGGCGG TGGTGGATCTGGAGGAGGGGGCAGCGGAGGTGGCGGCTCTGAG GTGAAACTGCAGGAGAGTGGCCCTGGCCTGGTGGCTCCCAGCCA GAGCCTTTCTGTCACCTGCACCGTGTCTGGGGTGTCCCTGCCTGA

[0235] HDRT17

[0236] 49 CTATGGAGTCTCCTGGATCCGGCAGCCTCCAAGAAAAGGACTGGA (Z2 / Z6) ATGGCTGGGCGTCATCTGGGGAAGTGAGACCACCTACTATAATTC AGCCCTCAAGTCCCGGCTCACCATCATTAAGGACAACTCCAAATCC CAGGTGTTCCTGAAGATGAATTCTCTCCAGACTGATGACACAGCCA TCTACTACTGTGCCAAGCATTATTATTACGGCGGGTCCTATGCCAT GGACTACTGGGGCCAGGGGACCAGTGTCACTGTTTCTTCTGAAAG CAAGTATGGCCCCCCATGCCCTCCCTGCCCCGGACAGCCGCGGG AGCCCCAGGTCTACACCCTCCCACCCTCCAGAGATGAGCTGACCA AGAACCAGGTTTCACTGACATGCCTGGTGAAGGGCTTCTACCCCT CTGACATTGCTGTGGAGTGGGAGAGCAATGGGCAGCCAGAGAACA ACTACAAGACCACACCTCCTGTCCTGGACAGCGACGGCTCCTTCTT CCTCTACAGTAAGCTCACTGTGGACAAGAGCCGCTGGCAGCAGGG CAATGTCTTCTCCTGCAGCGTGATGCACGAGGCCCTGCACAATGC CTATACCCAGAAGTCACTCTCTCTGAGCCCTGGAAAGAAGGATCCC AAGTTTTGGGTCCTCGTGGTGGTCGGGGGTGTGCTGGCCTGCTAC AGCTTGCTGGTCACAGTGGCCTTCATCATCTTCTGGGTGCGCTCCA AGAGGAGCCGGCTGCTTCACAGTGATTACATGAACATGACCCCCA GGAGGCCAGGACCCACCAGGAAGCACTACCAGCCCTACGCTCCC

[0237] CCGCGGGACTTTGCTGCTTACCGCAGCagagtgaaggtgggtaccactgggct ttgggaggagggcacggggtcccccacttgatggatgttcagaggggccttggtcttggaaggtctc aagctcgggtggtgcctggggcttggtatccaggagcaaagcaaggaccagccaagtgtgtgcct tgagtgggctgaggaggaggtggcagtgtctggctgagatggacagggtaggagggagagcct ggtgctaggcacctccatgacaagccgtacaaatgtgtgcacatcagagtgtcccagggaaggc gatgctactggtgacaaaggggcttacactcaggcagaggtccttctttccaagtgtgaatgaaggc catgttagcctttctcttgaaaaggccctttcctcatctgtaactggggaG _ tcttcccgaggcttccagagttcttggttcacaccggggctaacatggctgggcttctgctgcagtggc aggagctctgtgcacagagaacagcctcatctgctcgccttgtttccacctcccctcccattgcccca ggttctttggccccacagcggccacatctgccgttggtgccaataggttttccaggagctggttgagg tgggagggagggagagggttgtgatcaggctgaggcatggggattggatatagtctccgtgtcatg atttatttggtcagtcagtcctagtgccaccctggggtaatggggatgtgttctcgtcaccttgggcctg gctgaccagctttatctcttggcacagaggcacagagctttggcctgctggatcccaaaggcagcg gcGCTACCAATTTTTCTTTGTTGAAGCAGGCCGGGGATGTGGAGGA AAATCCGGGGCCAATGGCTCTTCCTGTGACTGCCCTTCTGCTGCC CCTGGCTCTCCTGCTCCATGCTGCCCGGCCAGACATCCAGATGAC ACAGACAACCAGCAGCCTCAGCGCCAGCCTGGGGGACAGAGTGA CCATTAGCTGCCGGGCCTCTCAGGACATCAGCAAATACCTGAACT GGTACCAGCAGAAACCAGATGGCACTGTCAAGCTGCTGATTTACC ACACCTCCAGGCTCCACAGCGGCGTGCCCAGTCGCTTCAGCGGC AGTGGGAGCGGGACAGATTATTCCCTCACAATCTCCAACCTGGAG CAGGAAGATATTGCCACATACTTCTGCCAGCAAGGCAACACCCTG CCATACACATTTGGAGGCGGCACCAAATTGGAGATCACCGGCGGT GGTGGATCTGGAGGAGGGGGCAGCGGAGGTGGCGGCTCTGAGG TGAAACTGCAGGAGAGTGGCCCTGGCCTGGTGGCTCCCAGCCAG AGCCTTTCTGTCACCTGCACCGTGTCTGGGGTGTCCCTGCCTGAC TATGGAGTCTCCTGGATCCGGCAGCCTCCAAGAAAAGGACTGGAA TGGCTGGGCGTCATCTGGGGAAGTGAGACCACCTACTATAATTCA GCCCTCAAGTCCCGGCTCACCATCATTAAGGACAACTCCAAATCCC HDRT20

[0238] 50 AGGTGTTCCTGAAGATGAATTCTCTCCAGACTGATGACACAGCCAT (Z3 / Z5) CTACTACTGTGCCAAGCATTATTATTACGGCGGGTCCTATGCCATG GACTACTGGGGCCAGGGGACCAGTGTCACTGTTTCTTCTGAAAGC AAGTATGGCCCCCCATGCCCTCCCTGCCCCGGACAGCCGCGGGA GCCCCAGGTCTACACCCTCCCACCCTCCAGAGATGAGCTGACCAA GAACCAGGTTTCACTGACATGCCTGGTGAAGGGCTTCTACCCCTCT GACATTGCTGTGGAGTGGGAGAGCAATGGGCAGCCAGAGAACAAC TACAAGACCACACCTCCTGTCCTGGACAGCGACGGCTCCTTCTTC CTCTACAGTAAGCTCACTGTGGACAAGAGCCGCTGGCAGCAGGGC AATGTCTTCTCCTGCAGCGTGATGCACGAGGCCCTGCACAATGCC TATACCCAGAAGTCACTCTCTCTGAGCCCTGGAAAGAAGGATCCCA AGTTTTGGGTCCTCGTGGTGGTCGGGGGTGTGCTGGCCTGCTACA GCTTGCTGGTCACAGTGGCCTTCATCATCTTCTGGGTGCGCTCCAA GAGGAGCCGGCTGCTTCACAGTGATTACATGAACATGACCCCCAG GAGGCCAGGACCCACCAGGAAGCACTACCAGCCCTACGCTCCCC CGCGGGACTTTGCTGCTTACCGCAGCagagtgaaggtgggtaccactgggcttt gggaggagggcacggggtcccccacttgatggatgttcagaggggccttggtcttggaaggtctca agctcgggtggtgcctggggcttggtatccaggagcaaagcaaggaccagccaagtgtgtgcctt gagtgggctgaggaggaggtggcagtgtctggctgagatggacagggtaggagggagagcctg gtgctaggcacctccatgacaagccgtacaaatgtgtgcacatcagagtgtcccagggaaggcg atgctactggtgacaaaggggcttacactcaggcagaggtccttctttccaagtgtgaatgaaggcc atgttagcctttctcttgaaaaggccctttcctcatctgtaactggggaG _ tcttcccgaggcttccagagttcttggttcacaccggggctaacatggctgggcttctgctgcagtggc aggagctctgtgcacagagaacagcctcatctgctcgccttgtttccacctcccctcccattgcccca ggttctttggccccacagcggccacatctgccgttggtgccaataggttttccaggagctggttgagg tgggagggagggagagggttgtgatcaggctgaggcatggggattggatatagtctccgtgtcatg atttatttggtcagtcagtcctagtgccaccctggggtaatggggatgtgttctcgtcaccttgggcctg HDRT20

[0239] 51 gctgaccagctttatctcttggcacagaggcacagagctttggcctgctggatcccaaaggcagcg (Z3 / Z6) gcGCTACCAATTTTTCTTTGTTGAAGCAGGCCGGGGATGTGGAGGA AAATCCGGGGCCAATGGCTCTTCCTGTGACTGCCCTTCTGCTGCC CCTGGCTCTCCTGCTCCATGCTGCCCGGCCAGACATCCAGATGAC ACAGACAACCAGCAGCCTCAGCGCCAGCCTGGGGGACAGAGTGA CCATTAGCTGCCGGGCCTCTCAGGACATCAGCAAATACCTGAACT GGTACCAGCAGAAACCAGATGGCACTGTCAAGCTGCTGATTTACC ACACCTCCAGGCTCCACAGCGGCGTGCCCAGTCGCTTCAGCGGC AGTGGGAGCGGGACAGATTATTCCCTCACAATCTCCAACCTGGAG CAGGAAGATATTGCCACATACTTCTGCCAGCAAGGCAACACCCTG CCATACACATTTGGAGGCGGCACCAAATTGGAGATCACCGGCGGT GGTGGATCTGGAGGAGGGGGCAGCGGAGGTGGCGGCTCTGAGG TGAAACTGCAGGAGAGTGGCCCTGGCCTGGTGGCTCCCAGCCAG AGCCTTTCTGTCACCTGCACCGTGTCTGGGGTGTCCCTGCCTGAC TATGGAGTCTCCTGGATCCGGCAGCCTCCAAGAAAAGGACTGGAA TGGCTGGGCGTCATCTGGGGAAGTGAGACCACCTACTATAATTCA GCCCTCAAGTCCCGGCTCACCATCATTAAGGACAACTCCAAATCCC AGGTGTTCCTGAAGATGAATTCTCTCCAGACTGATGACACAGCCAT CTACTACTGTGCCAAGCATTATTATTACGGCGGGTCCTATGCCATG GACTACTGGGGCCAGGGGACCAGTGTCACTGTTTCTTCTGAAAGC AAGTATGGCCCCCCATGCCCTCCCTGCCCCGGACAGCCGCGGGA GCCCCAGGTCTACACCCTCCCACCCTCCAGAGATGAGCTGACCAA GAACCAGGTTTCACTGACATGCCTGGTGAAGGGCTTCTACCCCTCT GACATTGCTGTGGAGTGGGAGAGCAATGGGCAGCCAGAGAACAAC TACAAGACCACACCTCCTGTCCTGGACAGCGACGGCTCCTTCTTC CTCTACAGTAAGCTCACTGTGGACAAGAGCCGCTGGCAGCAGGGC AATGTCTTCTCCTGCAGCGTGATGCACGAGGCCCTGCACAATGCC TATACCCAGAAGTCACTCTCTCTGAGCCCTGGAAAGAAGGATCCCA AGTTTTGGGTCCTCGTGGTGGTCGGGGGTGTGCTGGCCTGCTACA GCTTGCTGGTCACAGTGGCCTTCATCATCTTCTGGGTGCGCTCCAA GAGGAGCCGGCTGCTTCACAGTGATTACATGAACATGACCCCCAG GAGGCCAGGACCCACCAGGAAGCACTACCAGCCCTACGCTCCCC CGCGGGACTTTGCTGCTTACCGCAGCagagtgaaggtgggtaccactgggcttt gggaggagggcacggggtcccccacttgatggatgttcagaggggccttggtcttggaaggtctca agctcgggtggtgcctggggcttggtatccaggagcaaagcaaggaccagccaagtgtgtgcctt gagtgggctgaggaggaggtggcagtgtctggctgagatggacagggtaggagggagagcctg gtgctaggcacctccatgacaagccgtacaaatgtgtgcacatcagagtgtcccagggaaggcg atgctactggtgacaaaggggcttacactcaggcagaggtccttctttccaagtgtgaatgaaggcc atgttagcctttctcttgaaaaggccctttcctcatctgtaactggggaG _ tcttcccgaggcttccagagttcttggttcacaccggggctaacatggctgggcttctgctgcagtggc aggagctctgtgcacagagaacagcctcatctgctcgccttgtttccacctcccctcccattgcccca ggttctttggccccacagcggccacatctgccgttggtgccaataggttttccaggagctggttgagg tgggagggagggagagggttgtgatcaggctgaggcatggggattggatatagtctccgtgtcatg atttatttggtcagtcagtcctagtgccaccctggggtaatggggatgtgttctcgtcaccttgggcctg gctgaccagctttatctcttggcacagaggcacagagctttggcctgctggatcccaaaggcagcg gcGCTACCAATTTTTCTTTGTTGAAGCAGGCCGGGGATGTGGAGGA AAATCCGGGGCCAATGGCTCTTCCTGTGACTGCCCTTCTGCTGCC CCTGGCTCTCCTGCTCCATGCTGCCCGGCCAGACATCCAGATGAC ACAGACAACCAGCAGCCTCAGCGCCAGCCTGGGGGACAGAGTGA CCATTAGCTGCCGGGCCTCTCAGGACATCAGCAAATACCTGAACT GGTACCAGCAGAAACCAGATGGCACTGTCAAGCTGCTGATTTACC HDRT21

[0240] 52 ACACCTCCAGGCTCCACAGCGGCGTGCCCAGTCGCTTCAGCGGC (Z3 / Z7) AGTGGGAGCGGGACAGATTATTCCCTCACAATCTCCAACCTGGAG CAGGAAGATATTGCCACATACTTCTGCCAGCAAGGCAACACCCTG CCATACACATTTGGAGGCGGCACCAAATTGGAGATCACCGGCGGT GGTGGATCTGGAGGAGGGGGCAGCGGAGGTGGCGGCTCTGAGG TGAAACTGCAGGAGAGTGGCCCTGGCCTGGTGGCTCCCAGCCAG AGCCTTTCTGTCACCTGCACCGTGTCTGGGGTGTCCCTGCCTGAC TATGGAGTCTCCTGGATCCGGCAGCCTCCAAGAAAAGGACTGGAA TGGCTGGGCGTCATCTGGGGAAGTGAGACCACCTACTATAATTCA GCCCTCAAGTCCCGGCTCACCATCATTAAGGACAACTCCAAATCCC AGGTGTTCCTGAAGATGAATTCTCTCCAGACTGATGACACAGCCAT CTACTACTGTGCCAAGCATTATTATTACGGCGGGTCCTATGCCATG GACTACTGGGGCCAGGGGACCAGTGTCACTGTTTCTTCTGAAAGC AAGTATGGCCCCCCATGCCCTCCCTGCCCCGGACAGCCGCGGGA GCCCCAGGTCTACACCCTCCCACCCTCCAGAGATGAGCTGACCAA GAACCAGGTTTCACTGACATGCCTGGTGAAGGGCTTCTACCCCTCT GACATTGCTGTGGAGTGGGAGAGCAATGGGCAGCCAGAGAACAAC TACAAGACCACACCTCCTGTCCTGGACAGCGACGGCTCCTTCTTC CTCTACAGTAAGCTCACTGTGGACAAGAGCCGCTGGCAGCAGGGC AATGTCTTCTCCTGCAGCGTGATGCACGAGGCCCTGCACAATGCC TATACCCAGAAGTCACTCTCTCTGAGCCCTGGAAAGAAGGATCCCA AGTTTTGGGTCCTCGTGGTGGTCGGGGGTGTGCTGGCCTGCTACA GCTTGCTGGTCACAGTGGCCTTCATCATCTTCTGGGTGCGCTCCAA GAGGAGCCGGCTGCTTCACAGTGATTACATGAACATGACCCCCAG GAGGCCAGGACCCACCAGGAAGCACTACCAGCCCTACGCTCCCC CGCGGGACTTTGCTGCTTACCGCAGCagagtgaaggtgggtaccactgCgcttt gCgaggagggcacggggtcccccacttgatggatgttcagaggggccttggtcttggaaggtctc aagctcgggtggtgcctggggcttggtatccaggagcaaagcaaggaccagccaagtgtgtgcct tgagtgggctgaggaggaggtggcagtgtctggctgagatggacagggtaggagggagagcct ggtgctaggcacctccatgacaagccgtacaaatgtgtgcacatcagagtgtcccagggaaggc gatgctactggtgacaaaggggcttacactcaggcagaggtccttctttccaagtgtgaatgaaggc catgttagcctttctcttgaaaaggccctttcctcatctgtaactggggaG _ tcttcccgaggcttccagagttcttggttcacaccggggctaacatggctgggcttctgctgcagtggc aggagctctgtgcacagagaacagcctcatctgctcgccttgtttccacctcccctcccattgcccca ggttctttggccccacagcggccacatctgccgttggtgccaataggttttccaggagctggttgagg tgggagggagggagagggttgtgatcaggctgaggcatggggattggatatagtctccgtgtcatg atttatttggtcagtcagtcctagtgccaccctggggtaatggggatgtgttctcgtcaccttgggcctg gctgaccagctttatctcttggcacagaggcacagagctttggcctgctggatcccaaaggcagcg gcGCTACCAATTTTTCTTTGTTGAAGCAGGCCGGGGATGTGGAGGA AAATCCGGGGCCAATGGCTCTTCCTGTGACTGCCCTTCTGCTGCC CCTGGCTCTCCTGCTCCATGCTGCCCGGCCAGACATCCAGATGAC ACAGACAACCAGCAGCCTCAGCGCCAGCCTGGGGGACAGAGTGA CCATTAGCTGCCGGGCCTCTCAGGACATCAGCAAATACCTGAACT GGTACCAGCAGAAACCAGATGGCACTGTCAAGCTGCTGATTTACC ACACCTCCAGGCTCCACAGCGGCGTGCCCAGTCGCTTCAGCGGC AGTGGGAGCGGGACAGATTATTCCCTCACAATCTCCAACCTGGAG CAGGAAGATATTGCCACATACTTCTGCCAGCAAGGCAACACCCTG CCATACACATTTGGAGGCGGCACCAAATTGGAGATCACCGGCGGT GGTGGATCTGGAGGAGGGGGCAGCGGAGGTGGCGGCTCTGAGG TGAAACTGCAGGAGAGTGGCCCTGGCCTGGTGGCTCCCAGCCAG AGCCTTTCTGTCACCTGCACCGTGTCTGGGGTGTCCCTGCCTGAC HDRT21

[0241] 53 TATGGAGTCTCCTGGATCCGGCAGCCTCCAAGAAAAGGACTGGAA (Z3 / Z8) TGGCTGGGCGTCATCTGGGGAAGTGAGACCACCTACTATAATTCA GCCCTCAAGTCCCGGCTCACCATCATTAAGGACAACTCCAAATCCC AGGTGTTCCTGAAGATGAATTCTCTCCAGACTGATGACACAGCCAT CTACTACTGTGCCAAGCATTATTATTACGGCGGGTCCTATGCCATG GACTACTGGGGCCAGGGGACCAGTGTCACTGTTTCTTCTGAAAGC AAGTATGGCCCCCCATGCCCTCCCTGCCCCGGACAGCCGCGGGA GCCCCAGGTCTACACCCTCCCACCCTCCAGAGATGAGCTGACCAA GAACCAGGTTTCACTGACATGCCTGGTGAAGGGCTTCTACCCCTCT GACATTGCTGTGGAGTGGGAGAGCAATGGGCAGCCAGAGAACAAC TACAAGACCACACCTCCTGTCCTGGACAGCGACGGCTCCTTCTTC CTCTACAGTAAGCTCACTGTGGACAAGAGCCGCTGGCAGCAGGGC AATGTCTTCTCCTGCAGCGTGATGCACGAGGCCCTGCACAATGCC TATACCCAGAAGTCACTCTCTCTGAGCCCTGGAAAGAAGGATCCCA AGTTTTGGGTCCTCGTGGTGGTCGGGGGTGTGCTGGCCTGCTACA GCTTGCTGGTCACAGTGGCCTTCATCATCTTCTGGGTGCGCTCCAA GAGGAGCCGGCTGCTTCACAGTGATTACATGAACATGACCCCCAG GAGGCCAGGACCCACCAGGAAGCACTACCAGCCCTACGCTCCCC CGCGGGACTTTGCTGCTTACCGCAGCagagtgaaggtgggtaccactgCgcttt gCgaggagggcacggggtcccccacttgatggatgttcagaggggccttggtcttggaaggtctc aagctcgggtggtgcctggggcttggtatccaggagcaaagcaaggaccagccaagtgtgtgcct tgagtgggctgaggaggaggtggcagtgtctggctgagatggacagggtaggagggagagcct ggtgctaggcacctccatgacaagccgtacaaatgtgtgcacatcagagtgtcccagggaaggc gatgctactggtgacaaaggggcttacactcaggcagaggtccttctttccaagtgtgaatgaaggc catgttagcctttctcttgaaaaggccctttcctcatctgtaactggggaG _ acaccggggctaacatggctgggcttctgctgcagtggcaggagctctgtgcacagagaacagc ctcatctgctcgccttgtttccacctcccctcccattgccccaggttctttggccccacagcggccacat ctgccgttggtgccaataggttttccaggagctggttgaggtgggagggagggagagggttgtgatc aggctgaggcatggggattggatatagtctccgtgtcatgatttatttggtcagtcagtcctagtgcca ccctggggtaatggggatgtgttctcgtcaccttgggcctggctgaccagctttatctcttggcacaga ggcacagagctttggcctgctggatcTcaaactctgctacctgctggatggaatcctcttcggcagc ggcGCTACCAATTTTTCTTTGTTGAAGCAGGCCGGGGATGTGGAGG AAAATCCGGGGCCAATGGCTCTTCCTGTGACTGCCCTTCTGCTGC CCCTGGCTCTCCTGCTCCATGCTGCCCGGCCAGACATCCAGATGA CACAGACAACCAGCAGCCTCAGCGCCAGCCTGGGGGACAGAGTG ACCATTAGCTGCCGGGCCTCTCAGGACATCAGCAAATACCTGAACT GGTACCAGCAGAAACCAGATGGCACTGTCAAGCTGCTGATTTACC ACACCTCCAGGCTCCACAGCGGCGTGCCCAGTCGCTTCAGCGGC AGTGGGAGCGGGACAGATTATTCCCTCACAATCTCCAACCTGGAG CAGGAAGATATTGCCACATACTTCTGCCAGCAAGGCAACACCCTG CCATACACATTTGGAGGCGGCACCAAATTGGAGATCACCGGCGGT GGTGGATCTGGAGGAGGGGGCAGCGGAGGTGGCGGCTCTGAGG TGAAACTGCAGGAGAGTGGCCCTGGCCTGGTGGCTCCCAGCCAG AGCCTTTCTGTCACCTGCACCGTGTCTGGGGTGTCCCTGCCTGAC TATGGAGTCTCCTGGATCCGGCAGCCTCCAAGAAAAGGACTGGAA TGGCTGGGCGTCATCTGGGGAAGTGAGACCACCTACTATAATTCA GCCCTCAAGTCCCGGCTCACCATCATTAAGGACAACTCCAAATCCC HDRT22

[0242] 54 AGGTGTTCCTGAAGATGAATTCTCTCCAGACTGATGACACAGCCAT (Z3 / Z5) CTACTACTGTGCCAAGCATTATTATTACGGCGGGTCCTATGCCATG GACTACTGGGGCCAGGGGACCAGTGTCACTGTTTCTTCTGAAAGC AAGTATGGCCCCCCATGCCCTCCCTGCCCCGGACAGCCGCGGGA GCCCCAGGTCTACACCCTCCCACCCTCCAGAGATGAGCTGACCAA GAACCAGGTTTCACTGACATGCCTGGTGAAGGGCTTCTACCCCTCT GACATTGCTGTGGAGTGGGAGAGCAATGGGCAGCCAGAGAACAAC TACAAGACCACACCTCCTGTCCTGGACAGCGACGGCTCCTTCTTC CTCTACAGTAAGCTCACTGTGGACAAGAGCCGCTGGCAGCAGGGC AATGTCTTCTCCTGCAGCGTGATGCACGAGGCCCTGCACAATGCC TATACCCAGAAGTCACTCTCTCTGAGCCCTGGAAAGAAGGATCCCA AGTTTTGGGTCCTCGTGGTGGTCGGGGGTGTGCTGGCCTGCTACA GCTTGCTGGTCACAGTGGCCTTCATCATCTTCTGGGTGCGCTCCAA GAGGAGCCGGCTGCTTCACAGTGATTACATGAACATGACCCCCAG GAGGCCAGGACCCACCAGGAAGCACTACCAGCCCTACGCTCCCC CGCGGGACTTTGCTGCTTACCGCAGCagagtgaaggtgggtaccactgggcttt gggaggagggcacggggtcccccacttgatggatgttcagaggggccttggtcttggaaggtctca agctcgggtggtgcctggggcttggtatccaggagcaaagcaaggaccagccaagtgtgtgcctt gagtgggctgaggaggaggtggcagtgtctggctgagatggacagggtaggagggagagcctg gtgctaggcacctccatgacaagccgtacaaatgtgtgcacatcagagtgtcccagggaaggcg atgctactggtgacaaaggggcttacactcaggcagaggtccttctttccaagtgtgaatgaaggcc atgttagcctttctcttgaaaaggccctttcctcatctgtaactggggaG _ acaccggggctaacatggctgggcttctgctgcagtggcaggagctctgtgcacagagaacagc ctcatctgctcgccttgtttccacctcccctcccattgccccaggttctttggccccacagcggccacat ctgccgttggtgccaataggttttccaggagctggttgaggtgggagggagggagagggttgtgatc aggctgaggcatggggattggatatagtctccgtgtcatgatttatttggtcagtcagtcctagtgcca HDRT22

[0243] 55 ccctggggtaatggggatgtgttctcgtcaccttgggcctggctgaccagctttatctcttggcacaga (Z3 / Z6) ggcacagagctttggcctgctggatcTcaaactctgctacctgctggatggaatcctcttcggcagc ggcGCTACCAATTTTTCTTTGTTGAAGCAGGCCGGGGATGTGGAGG AAAATCCGGGGCCAATGGCTCTTCCTGTGACTGCCCTTCTGCTGC CCCTGGCTCTCCTGCTCCATGCTGCCCGGCCAGACATCCAGATGA CACAGACAACCAGCAGCCTCAGCGCCAGCCTGGGGGACAGAGTG ACCATTAGCTGCCGGGCCTCTCAGGACATCAGCAAATACCTGAACT GGTACCAGCAGAAACCAGATGGCACTGTCAAGCTGCTGATTTACC ACACCTCCAGGCTCCACAGCGGCGTGCCCAGTCGCTTCAGCGGC AGTGGGAGCGGGACAGATTATTCCCTCACAATCTCCAACCTGGAG CAGGAAGATATTGCCACATACTTCTGCCAGCAAGGCAACACCCTG CCATACACATTTGGAGGCGGCACCAAATTGGAGATCACCGGCGGT GGTGGATCTGGAGGAGGGGGCAGCGGAGGTGGCGGCTCTGAGG TGAAACTGCAGGAGAGTGGCCCTGGCCTGGTGGCTCCCAGCCAG AGCCTTTCTGTCACCTGCACCGTGTCTGGGGTGTCCCTGCCTGAC TATGGAGTCTCCTGGATCCGGCAGCCTCCAAGAAAAGGACTGGAA TGGCTGGGCGTCATCTGGGGAAGTGAGACCACCTACTATAATTCA GCCCTCAAGTCCCGGCTCACCATCATTAAGGACAACTCCAAATCCC AGGTGTTCCTGAAGATGAATTCTCTCCAGACTGATGACACAGCCAT CTACTACTGTGCCAAGCATTATTATTACGGCGGGTCCTATGCCATG GACTACTGGGGCCAGGGGACCAGTGTCACTGTTTCTTCTGAAAGC AAGTATGGCCCCCCATGCCCTCCCTGCCCCGGACAGCCGCGGGA GCCCCAGGTCTACACCCTCCCACCCTCCAGAGATGAGCTGACCAA GAACCAGGTTTCACTGACATGCCTGGTGAAGGGCTTCTACCCCTCT GACATTGCTGTGGAGTGGGAGAGCAATGGGCAGCCAGAGAACAAC TACAAGACCACACCTCCTGTCCTGGACAGCGACGGCTCCTTCTTC CTCTACAGTAAGCTCACTGTGGACAAGAGCCGCTGGCAGCAGGGC AATGTCTTCTCCTGCAGCGTGATGCACGAGGCCCTGCACAATGCC TATACCCAGAAGTCACTCTCTCTGAGCCCTGGAAAGAAGGATCCCA AGTTTTGGGTCCTCGTGGTGGTCGGGGGTGTGCTGGCCTGCTACA GCTTGCTGGTCACAGTGGCCTTCATCATCTTCTGGGTGCGCTCCAA GAGGAGCCGGCTGCTTCACAGTGATTACATGAACATGACCCCCAG GAGGCCAGGACCCACCAGGAAGCACTACCAGCCCTACGCTCCCC CGCGGGACTTTGCTGCTTACCGCAGCagagtgaaggtgggtaccactgggcttt gggaggagggcacggggtcccccacttgatggatgttcagaggggccttggtcttggaaggtctca agctcgggtggtgcctggggcttggtatccaggagcaaagcaaggaccagccaagtgtgtgcctt gagtgggctgaggaggaggtggcagtgtctggctgagatggacagggtaggagggagagcctg gtgctaggcacctccatgacaagccgtacaaatgtgtgcacatcagagtgtcccagggaaggcg atgctactggtgacaaaggggcttacactcaggcagaggtccttctttccaagtgtgaatgaaggcc atgttagcctttctcttgaaaaggccctttcctcatctgtaactggggaG _ atctcatttgaccatcatttgacctattgtctccctgtgggtaggcctcagagccacacacctcaggcc aggagtaccattcatccagacgtgaacatcttcccgaggcttccagagttcttggttcacaccgggg ctaacatggctgggcttctgctgcagtggcaggagctctgtgcacagagaacagcctcatctgctcg ccttgtttccacctcccctcccattgccccaggttctttggccccacagcggccacatctgccgttggtg ccaataggttttccaggagctggttgaggtgggagggagggagagggttgtgatcaggctgaggc atggggattggatatagtctccgtgtcatgatttatttggtcagtcagtcctagtgcGaccctggggta atggggatgtgttctcgtcacGttgggcctggctgaccagctttatctcttggcacagaggcaggca gcggcGCTACCAATTTTTCTTTGTTGAAGCAGGCCGGGGATGTGGAG GAAAATCCGGGGCCAATGGCTCTTCCTGTGACTGCCCTTCTGCTG CCCCTGGCTCTCCTGCTCCATGCTGCCCGGCCAGACATCCAGATG ACACAGACAACCAGCAGCCTCAGCGCCAGCCTGGGGGACAGAGT HDRT23

[0244] 56 GACCATTAGCTGCCGGGCCTCTCAGGACATCAGCAAATACCTGAA (Z3 / Z7) CTGGTACCAGCAGAAACCAGATGGCACTGTCAAGCTGCTGATTTAC CACACCTCCAGGCTCCACAGCGGCGTGCCCAGTCGCTTCAGCGG CAGTGGGAGCGGGACAGATTATTCCCTCACAATCTCCAACCTGGA GCAGGAAGATATTGCCACATACTTCTGCCAGCAAGGCAACACCCT GCCATACACATTTGGAGGCGGCACCAAATTGGAGATCACCGGCGG TGGTGGATCTGGAGGAGGGGGCAGCGGAGGTGGCGGCTCTGAG GTGAAACTGCAGGAGAGTGGCCCTGGCCTGGTGGCTCCCAGCCA GAGCCTTTCTGTCACCTGCACCGTGTCTGGGGTGTCCCTGCCTGA CTATGGAGTCTCCTGGATCCGGCAGCCTCCAAGAAAAGGACTGGA ATGGCTGGGCGTCATCTGGGGAAGTGAGACCACCTACTATAATTC AGCCCTCAAGTCCCGGCTCACCATCATTAAGGACAACTCCAAATCC CAGGTGTTCCTGAAGATGAATTCTCTCCAGACTGATGACACAGCCA TCTACTACTGTGCCAAGCATTATTATTACGGCGGGTCCTATGCCAT GGACTACTGGGGCCAGGGGACCAGTGTCACTGTTTCTTCTGAAAG CAAGTATGGCCCCCCATGCCCTCCCTGCCCCGGACAGCCGCGGG AGCCCCAGGTCTACACCCTCCCACCCTCCAGAGATGAGCTGACCA AGAACCAGGTTTCACTGACATGCCTGGTGAAGGGCTTCTACCCCT CTGACATTGCTGTGGAGTGGGAGAGCAATGGGCAGCCAGAGAACA ACTACAAGACCACACCTCCTGTCCTGGACAGCGACGGCTCCTTCTT CCTCTACAGTAAGCTCACTGTGGACAAGAGCCGCTGGCAGCAGGG CAATGTCTTCTCCTGCAGCGTGATGCACGAGGCCCTGCACAATGC CTATACCCAGAAGTCACTCTCTCTGAGCCCTGGAAAGAAGGATCCC AAGTTTTGGGTCCTCGTGGTGGTCGGGGGTGTGCTGGCCTGCTAC AGCTTGCTGGTCACAGTGGCCTTCATCATCTTCTGGGTGCGCTCCA AGAGGAGCCGGCTGCTTCACAGTGATTACATGAACATGACCCCCA GGAGGCCAGGACCCACCAGGAAGCACTACCAGCCCTACGCTCCC CCGCGGGACTTTGCTGCTTACCGCAGCagagtgaaggtgggtaccactgCgc tttgCgaggagggcacggggtcccccacttgatggatgttcagaggggccttggtcttggaaggtct caagctcgggtggtgcctggggcttggtatccaggagcaaagcaaggaccagccaagtgtgtgc cttgagtgggctgaggaggaggtggcagtgtctggctgagatggacagggtaggagggagagc ctggtgctaggcacctccatgacaagccgtacaaatgtgtgcacatcagagtgtcccagggaagg cgatgctactggtgacaaaggggcttacactcaggcagaggtccttctttccaagtgtgaatgaagg ccatgttagcctttctcttgaaaaggccctttcctcatctgtaactggggaG _ atctcatttgaccatcatttgacctattgtctccctgtgggtaggcctcagagccacacacctcaggcc aggagtaccattcatccagacgtgaacatcttcccgaggcttccagagttcttggttcacaccgggg ctaacatggctgggcttctgctgcagtggcaggagctctgtgcacagagaacagcctcatctgctcg ccttgtttccacctcccctcccattgccccaggttctttggccccacagcggccacatctgccgttggtg ccaataggttttccaggagctggttgaggtgggagggagggagagggttgtgatcaggctgaggc atggggattggatatagtctccgtgtcatgatttatttggtcagtcagtcctagtgcGaccctggggta atggggatgtgttctcgtcacGttgggcctggctgaccagctttatctcttggcacagaggcaggca gcggcGCTACCAATTTTTCTTTGTTGAAGCAGGCCGGGGATGTGGAG GAAAATCCGGGGCCAATGGCTCTTCCTGTGACTGCCCTTCTGCTG CCCCTGGCTCTCCTGCTCCATGCTGCCCGGCCAGACATCCAGATG ACACAGACAACCAGCAGCCTCAGCGCCAGCCTGGGGGACAGAGT GACCATTAGCTGCCGGGCCTCTCAGGACATCAGCAAATACCTGAA CTGGTACCAGCAGAAACCAGATGGCACTGTCAAGCTGCTGATTTAC CACACCTCCAGGCTCCACAGCGGCGTGCCCAGTCGCTTCAGCGG CAGTGGGAGCGGGACAGATTATTCCCTCACAATCTCCAACCTGGA GCAGGAAGATATTGCCACATACTTCTGCCAGCAAGGCAACACCCT GCCATACACATTTGGAGGCGGCACCAAATTGGAGATCACCGGCGG

[0245] HDRT23

[0246] 57 TGGTGGATCTGGAGGAGGGGGCAGCGGAGGTGGCGGCTCTGAG (Z3 / Z8) GTGAAACTGCAGGAGAGTGGCCCTGGCCTGGTGGCTCCCAGCCA GAGCCTTTCTGTCACCTGCACCGTGTCTGGGGTGTCCCTGCCTGA CTATGGAGTCTCCTGGATCCGGCAGCCTCCAAGAAAAGGACTGGA ATGGCTGGGCGTCATCTGGGGAAGTGAGACCACCTACTATAATTC AGCCCTCAAGTCCCGGCTCACCATCATTAAGGACAACTCCAAATCC CAGGTGTTCCTGAAGATGAATTCTCTCCAGACTGATGACACAGCCA TCTACTACTGTGCCAAGCATTATTATTACGGCGGGTCCTATGCCAT GGACTACTGGGGCCAGGGGACCAGTGTCACTGTTTCTTCTGAAAG CAAGTATGGCCCCCCATGCCCTCCCTGCCCCGGACAGCCGCGGG AGCCCCAGGTCTACACCCTCCCACCCTCCAGAGATGAGCTGACCA AGAACCAGGTTTCACTGACATGCCTGGTGAAGGGCTTCTACCCCT

[0247] CTGACATTGCTGTGGAGTGGGAGAGCAATGGGCAGCCAGAGAACA ACTACAAGACCACACCTCCTGTCCTGGACAGCGACGGCTCCTTCTT CCTCTACAGTAAGCTCACTGTGGACAAGAGCCGCTGGCAGCAGGG CAATGTCTTCTCCTGCAGCGTGATGCACGAGGCCCTGCACAATGC CTATACCCAGAAGTCACTCTCTCTGAGCCCTGGAAAGAAGGATCCC AAGTTTTGGGTCCTCGTGGTGGTCGGGGGTGTGCTGGCCTGCTAC AGCTTGCTGGTCACAGTGGCCTTCATCATCTTCTGGGTGCGCTCCA AGAGGAGCCGGCTGCTTCACAGTGATTACATGAACATGACCCCCA GGAGGCCAGGACCCACCAGGAAGCACTACCAGCCCTACGCTCCC CCGCGGGACTTTGCTGCTTACCGCAGCagagtgaaggtgggtaccactgCgc tttgCgaggagggcacggggtcccccacttgatggatgttcagaggggccttggtcttggaaggtct caagctcgggtggtgcctggggcttggtatccaggagcaaagcaaggaccagccaagtgtgtgc cttgagtgggctgaggaggaggtggcagtgtctggctgagatggacagggtaggagggagagc ctggtgctaggcacctccatgacaagccgtacaaatgtgtgcacatcagagtgtcccagggaagg cgatgctactggtgacaaaggggcttacactcaggcagaggtccttctttccaagtgtgaatgaagg ccatgttagcctttctcttgaaaaggccctttcctcatctgtaactggggaG _ acaccggggctaacatggctgggcttctgctgcagtggcaggagctctgtgcacagagaacagc ctcatctgctcgccttgtttccacctcccctcccattgccccaggttctttggccccacagcggccacat ctgccgttggtgccaataggttttccaggagctggttgaggtgggagggagggagagggttgtgatc aggctgaggcatggggattggatatagtctccgtgtcatgatttatttggtcagtcagtcctagtgcca ccctggggtaatggggatgtgttctcgtcaccttgggcctggctgaccagctttatctcttggcacaga ggcacagagctttggcctgctggatcccaaactctgctacctgctggatggaatcctcttcggcagc ggcGCTACCAATTTTTCTTTGTTGAAGCAGGCCGGGGATGTGGAGG AAAATCCGGGGCCAATGGCTCTTCCTGTGACTGCCCTTCTGCTGC CCCTGGCTCTCCTGCTCCATGCTGCCCGGCCAGACATCCAGATGA CACAGACAACCAGCAGCCTCAGCGCCAGCCTGGGGGACAGAGTG ACCATTAGCTGCCGGGCCTCTCAGGACATCAGCAAATACCTGAACT GGTACCAGCAGAAACCAGATGGCACTGTCAAGCTGCTGATTTACC ACACCTCCAGGCTCCACAGCGGCGTGCCCAGTCGCTTCAGCGGC AGTGGGAGCGGGACAGATTATTCCCTCACAATCTCCAACCTGGAG CAGGAAGATATTGCCACATACTTCTGCCAGCAAGGCAACACCCTG CCATACACATTTGGAGGCGGCACCAAATTGGAGATCACCGGCGGT GGTGGATCTGGAGGAGGGGGCAGCGGAGGTGGCGGCTCTGAGG TGAAACTGCAGGAGAGTGGCCCTGGCCTGGTGGCTCCCAGCCAG AGCCTTTCTGTCACCTGCACCGTGTCTGGGGTGTCCCTGCCTGAC TATGGAGTCTCCTGGATCCGGCAGCCTCCAAGAAAAGGACTGGAA TGGCTGGGCGTCATCTGGGGAAGTGAGACCACCTACTATAATTCA GCCCTCAAGTCCCGGCTCACCATCATTAAGGACAACTCCAAATCCC HDRT24

[0248] 58 AGGTGTTCCTGAAGATGAATTCTCTCCAGACTGATGACACAGCCAT (Z4 / Z5) CTACTACTGTGCCAAGCATTATTATTACGGCGGGTCCTATGCCATG GACTACTGGGGCCAGGGGACCAGTGTCACTGTTTCTTCTGAAAGC AAGTATGGCCCCCCATGCCCTCCCTGCCCCGGACAGCCGCGGGA GCCCCAGGTCTACACCCTCCCACCCTCCAGAGATGAGCTGACCAA GAACCAGGTTTCACTGACATGCCTGGTGAAGGGCTTCTACCCCTCT GACATTGCTGTGGAGTGGGAGAGCAATGGGCAGCCAGAGAACAAC TACAAGACCACACCTCCTGTCCTGGACAGCGACGGCTCCTTCTTC CTCTACAGTAAGCTCACTGTGGACAAGAGCCGCTGGCAGCAGGGC AATGTCTTCTCCTGCAGCGTGATGCACGAGGCCCTGCACAATGCC TATACCCAGAAGTCACTCTCTCTGAGCCCTGGAAAGAAGGATCCCA AGTTTTGGGTCCTCGTGGTGGTCGGGGGTGTGCTGGCCTGCTACA GCTTGCTGGTCACAGTGGCCTTCATCATCTTCTGGGTGCGCTCCAA GAGGAGCCGGCTGCTTCACAGTGATTACATGAACATGACCCCCAG GAGGCCAGGACCCACCAGGAAGCACTACCAGCCCTACGCTCCCC CGCGGGACTTTGCTGCTTACCGCAGCagagtgaaggtgggtaccactgggcttt gggaggagggcacggggtcccccacttgatggatgttcagaggggccttggtcttggaaggtctca agctcgggtggtgcctggggcttggtatccaggagcaaagcaaggaccagccaagtgtgtgcctt gagtgggctgaggaggaggtggcagtgtctggctgagatggacagggtaggagggagagcctg gtgctaggcacctccatgacaagccgtacaaatgtgtgcacatcagagtgtcccagggaaggcg atgctactggtgacaaaggggcttacactcaggcagaggtccttctttccaagtgtgaatgaaggcc atgttagcctttctcttgaaaaggccctttcctcatctgtaactggggaG _ atctcatttgaccatcatttgacctattgtctccctgtgggtaggcctcagagccacacacctcaggcc aggagtaccattcatccagacgtgaacatcttcccgaggcttccagagttcttggttcacaccgggg HDRT23

[0249] 59 ctaacatggctgggcttctgctgcagtggcaggagctctgtgcacagagaacagcctcatctgctcg (Z4 / Z7) ccttgtttccacctcccctcccattgccccaggttctttggccccacagcggccacatctgccgttggtg ccaataggttttccaggagctggttgaggtgggagggagggagagggttgtgatcaggctgaggc atggggattggatatagtctccgtgtcatgatttatttggtcagtcagtcctagtgcGaccctggggta atggggatgtgttctcgtcacGttgggcctggctgaccagctttatctcttggcacagaggcaggca gcggcGCTACCAATTTTTCTTTGTTGAAGCAGGCCGGGGATGTGGAG GAAAATCCGGGGCCAATGGCTCTTCCTGTGACTGCCCTTCTGCTG CCCCTGGCTCTCCTGCTCCATGCTGCCCGGCCAGACATCCAGATG ACACAGACAACCAGCAGCCTCAGCGCCAGCCTGGGGGACAGAGT GACCATTAGCTGCCGGGCCTCTCAGGACATCAGCAAATACCTGAA CTGGTACCAGCAGAAACCAGATGGCACTGTCAAGCTGCTGATTTAC CACACCTCCAGGCTCCACAGCGGCGTGCCCAGTCGCTTCAGCGG CAGTGGGAGCGGGACAGATTATTCCCTCACAATCTCCAACCTGGA GCAGGAAGATATTGCCACATACTTCTGCCAGCAAGGCAACACCCT GCCATACACATTTGGAGGCGGCACCAAATTGGAGATCACCGGCGG TGGTGGATCTGGAGGAGGGGGCAGCGGAGGTGGCGGCTCTGAG GTGAAACTGCAGGAGAGTGGCCCTGGCCTGGTGGCTCCCAGCCA GAGCCTTTCTGTCACCTGCACCGTGTCTGGGGTGTCCCTGCCTGA CTATGGAGTCTCCTGGATCCGGCAGCCTCCAAGAAAAGGACTGGA ATGGCTGGGCGTCATCTGGGGAAGTGAGACCACCTACTATAATTC AGCCCTCAAGTCCCGGCTCACCATCATTAAGGACAACTCCAAATCC CAGGTGTTCCTGAAGATGAATTCTCTCCAGACTGATGACACAGCCA TCTACTACTGTGCCAAGCATTATTATTACGGCGGGTCCTATGCCAT GGACTACTGGGGCCAGGGGACCAGTGTCACTGTTTCTTCTGAAAG CAAGTATGGCCCCCCATGCCCTCCCTGCCCCGGACAGCCGCGGG AGCCCCAGGTCTACACCCTCCCACCCTCCAGAGATGAGCTGACCA AGAACCAGGTTTCACTGACATGCCTGGTGAAGGGCTTCTACCCCT CTGACATTGCTGTGGAGTGGGAGAGCAATGGGCAGCCAGAGAACA ACTACAAGACCACACCTCCTGTCCTGGACAGCGACGGCTCCTTCTT CCTCTACAGTAAGCTCACTGTGGACAAGAGCCGCTGGCAGCAGGG CAATGTCTTCTCCTGCAGCGTGATGCACGAGGCCCTGCACAATGC CTATACCCAGAAGTCACTCTCTCTGAGCCCTGGAAAGAAGGATCCC AAGTTTTGGGTCCTCGTGGTGGTCGGGGGTGTGCTGGCCTGCTAC AGCTTGCTGGTCACAGTGGCCTTCATCATCTTCTGGGTGCGCTCCA AGAGGAGCCGGCTGCTTCACAGTGATTACATGAACATGACCCCCA GGAGGCCAGGACCCACCAGGAAGCACTACCAGCCCTACGCTCCC CCGCGGGACTTTGCTGCTTACCGCAGCagagtgaaggtgggtaccactgCgc tttgCgaggagggcacggggtcccccacttgatggatgttcagaggggccttggtcttggaaggtct caagctcgggtggtgcctggggcttggtatccaggagcaaagcaaggaccagccaagtgtgtgc cttgagtgggctgaggaggaggtggcagtgtctggctgagatggacagggtaggagggagagc ctggtgctaggcacctccatgacaagccgtacaaatgtgtgcacatcagagtgtcccagggaagg cgatgctactggtgacaaaggggcttacactcaggcagaggtccttctttccaagtgtgaatgaagg ccatgttagcctttctcttgaaaaggccctttcctcatctgtaactggggaG _ atctcatttgaccatcatttgacctattgtctccctgtgggtaggcctcagagccacacacctcaggcc aggagtaccattcatccagacgtgaacatcttcccgaggcttccagagttcttggttcacaccgggg ctaacatggctgggcttctgctgcagtggcaggagctctgtgcacagagaacagcctcatctgctcg ccttgtttccacctcccctcccattgccccaggttctttggccccacagcggccacatctgccgttggtg ccaataggttttccaggagctggttgaggtgggagggagggagagggttgtgatcaggctgaggc atggggattggatatagtctccgtgtcatgatttatttggtcagtcagtcctagtgcGaccctggggta atggggatgtgttctcgtcacGttgggcctggctgaccagctttatctcttggcacagaggcaggca gcggcGCTACCAATTTTTCTTTGTTGAAGCAGGCCGGGGATGTGGAG

[0250] HDRT23

[0251] 60 GAAAATCCGGGGCCAATGGCTCTTCCTGTGACTGCCCTTCTGCTG (Z4 / Z8) CCCCTGGCTCTCCTGCTCCATGCTGCCCGGCCAGACATCCAGATG ACACAGACAACCAGCAGCCTCAGCGCCAGCCTGGGGGACAGAGT GACCATTAGCTGCCGGGCCTCTCAGGACATCAGCAAATACCTGAA CTGGTACCAGCAGAAACCAGATGGCACTGTCAAGCTGCTGATTTAC CACACCTCCAGGCTCCACAGCGGCGTGCCCAGTCGCTTCAGCGG CAGTGGGAGCGGGACAGATTATTCCCTCACAATCTCCAACCTGGA GCAGGAAGATATTGCCACATACTTCTGCCAGCAAGGCAACACCCT GCCATACACATTTGGAGGCGGCACCAAATTGGAGATCACCGGCGG TGGTGGATCTGGAGGAGGGGGCAGCGGAGGTGGCGGCTCTGAG GTGAAACTGCAGGAGAGTGGCCCTGGCCTGGTGGCTCCCAGCCA GAGCCTTTCTGTCACCTGCACCGTGTCTGGGGTGTCCCTGCCTGA CTATGGAGTCTCCTGGATCCGGCAGCCTCCAAGAAAAGGACTGGA ATGGCTGGGCGTCATCTGGGGAAGTGAGACCACCTACTATAATTC AGCCCTCAAGTCCCGGCTCACCATCATTAAGGACAACTCCAAATCC CAGGTGTTCCTGAAGATGAATTCTCTCCAGACTGATGACACAGCCA TCTACTACTGTGCCAAGCATTATTATTACGGCGGGTCCTATGCCAT GGACTACTGGGGCCAGGGGACCAGTGTCACTGTTTCTTCTGAAAG CAAGTATGGCCCCCCATGCCCTCCCTGCCCCGGACAGCCGCGGG AGCCCCAGGTCTACACCCTCCCACCCTCCAGAGATGAGCTGACCA AGAACCAGGTTTCACTGACATGCCTGGTGAAGGGCTTCTACCCCT CTGACATTGCTGTGGAGTGGGAGAGCAATGGGCAGCCAGAGAACA ACTACAAGACCACACCTCCTGTCCTGGACAGCGACGGCTCCTTCTT CCTCTACAGTAAGCTCACTGTGGACAAGAGCCGCTGGCAGCAGGG CAATGTCTTCTCCTGCAGCGTGATGCACGAGGCCCTGCACAATGC CTATACCCAGAAGTCACTCTCTCTGAGCCCTGGAAAGAAGGATCCC AAGTTTTGGGTCCTCGTGGTGGTCGGGGGTGTGCTGGCCTGCTAC AGCTTGCTGGTCACAGTGGCCTTCATCATCTTCTGGGTGCGCTCCA AGAGGAGCCGGCTGCTTCACAGTGATTACATGAACATGACCCCCA GGAGGCCAGGACCCACCAGGAAGCACTACCAGCCCTACGCTCCC CCGCGGGACTTTGCTGCTTACCGCAGCagagtgaaggtgggtaccactgCgc tttgCgaggagggcacggggtcccccacttgatggatgttcagaggggccttggtcttggaaggtct caagctcgggtggtgcctggggcttggtatccaggagcaaagcaaggaccagccaagtgtgtgc cttgagtgggctgaggaggaggtggcagtgtctggctgagatggacagggtaggagggagagc ctggtgctaggcacctccatgacaagccgtacaaatgtgtgcacatcagagtgtcccagggaagg cgatgctactggtgacaaaggggcttacactcaggcagaggtccttctttccaagtgtgaatgaagg ccatgttagcctttctcttgaaaaggccctttcctcatctgtaactggggaG _ aagctcatttggccagagtggaaatggaattgggagaaatcgatgaccaaatgtaaacacttggtg cctgatatagcttgacaccaagttagccccaagtgaaataccctggcaatattaatgtgtcttttcccg atattcctcaggtactccaaagattcaggtttactcacgtcatccagcagagaatggaaagtcaaatt tcctgaattgctatgtgtctgggtttcatccatccgacattgaagttgacttactgaagaatggagaga gaattgaaaaagtggagcattcagacttgtctttcagcaaggactggtctttctatctcttgtactacact gaattcacccccactgaaaaagatgagtatgctTGCAGGGTTAACCATGTCACATT GTCACAACCTAAAATTGTGAAATGGGACAGGGACATGGGTGGGGG CGGATCAGGCGGCGGCGGTTCTGGGGGAGGCGGCTCAGGTGGC GGTGGGAGTGGATCCCATTCCCTGAAGTATTTCCACACCAGCGTTA GTCGGCCGGGAAGGGGAGAACCAAGATTCATTTCCGTCGGCTATG TCGACGATACCCAATTTGTGCGATTTGATAATGACGCAGCTTCACC

[0252] CCGCATGGTGCCTCGGGCTCCTTGGATGGAGCAAGAAGGCTCAGA GTACTGGGACCGGGAGACCCGATCTGCGCGCGATACAGCACAAAT CTTTAGGGTCAACCTTCGAACATTGAGGGGCTACTACAACCAGAGT GAGGCAGGTTCCCATACGTTGCAATGGATGCATGGTTGCGAACTT HDRT26

[0253] 61 GGTCCAGATGGGAGGTTCCTCAGAGGTTATGAACAATTTGCTTACG (B1 / B8) ATGGAAAGGACTACCTTACACTCAATGAGGACCTCCGCAGCTGGA CCGCCGTTGACACCGCTGCTCAAATCTCCGAGCAAAAGAGTAATG ACGCGTCAGAAGCGGAGCATCAGCGCGCCTACCTCGAGGACACG TGCGTTGAGTGGCTCCATAAATACCTGGAAAAAGGGAAAGAAACC CTCCTCCACCTGGAGCCACCGAAGACCCACGTCACGCATCACCCA ATATCCGATCACGAGGCTACACTGAGGTGCTGGGCTCTCGGTTTC TATCCGGCAGAAATAACCCTGACCTGGCAGCAGGATGGTGAGGGC CATACGCAGGACACGGAGCTTGTCGAGACTCGACCCGCCGGTGAT GGAACATTCCAAAAATGGGCTGCTGTCGTTGTACCCTCTGGAGAA GAACAACGATACACCTGCCACGTGCAGCATGAGGGACTGCCGGAA CCGGTAACCTTGCGCTGGAAGCCGGCGTCACAACCAACGATACCC ATCGTCGGCATTATAGCCGGACTTGTTCTCCTGGGAAGCGTGGTC AGCGGAGCTGTCGTAGCGGCCGTAATTTGGAGGAAAAAGTCTTCT GGGGGAAAGGGTGGGTCATACTCTAAAGCGGAGTGGTCTGACTCA GCACAGGGTTCCGAATCCCACTCCCTGTGAGCGGCCGCGTCGAGT CTAGAGGGCCCGTTTAAACCCGCTGATCAGCCTCGACTGTGCCTT CTAGTTGCCAGCCATCTGTTGTTTGCCCCTCCCCCGTGCCTTCCTT GACCCTGGAAGGTGCCACTCCCACTGTCCTTTCCTAATAAAATGAG GAAATTGCATCGCATTGTCTGAGTAGGTGTCATTCTATTCTGGGGG GTGGGGTGGGGCAGGACAGCAAGGGGGAGGATTGGGAAGACAAT AGCAGGCAT GCT GGGGAT GCGGT GGGCT CT AT GGccaagatagttaagt ggggtaagtcttacattcttttgtaagctgctgaaagttgtgtatgagtagtcatatcataaagctgctttg atataaaaaaggtctatggccatactaccctgaatgagtcccatcccatctgatataaacaatctgc atattgggattgtcagggaatgttcttaaagatcagattagtggcacctgctgagatactgatgcaca gcatggtttctgaaccagtagtttccctgcagttgagcagggagcagcagcagcacttgcacaaat acatatacactcttaacacttcttacctactggcttcctctagcttttgtggcagcttcaggta _ aagctcatttggccagagtggaaatggaattgggagaaatcgatgaccaaatgtaaacacttggtg cctgatatagcttgacaccaagttagccccaagtgaaataccctggcaatattaatgtgtcttttcccg atattcctcaggtactccaaagattcaggtttactcacgtcatccagcagagaatggaaagtcaaatt tcctgaattgctatgtgtctgggtttcatccatccgacattgaagttgacttactgaagaatggagaga gaattgaaaaagtggagcattcagacttgtctttcagcaaggactggtctttctatctcttgtactacact gaattcacccccactgaaaaagatgagtatgcCT GCCGT GT GAACCAT GT GACTT TGTCACAGCCCAAGATTGTGAAATGGGACAGGGACATGGGTGGGG GCGGATCAGGCGGCGGCGGTTCTGGGGGAGGCGGCTCAGGTGG CGGTGGGAGTGGATCCCATTCCCTGAAGTATTTCCACACCAGCGT TAGTCGGCCGGGAAGGGGAGAACCAAGATTCATTTCCGTCGGCTA TGTCGACGATACCCAATTTGTGCGATTTGATAATGACGCAGCTTCA CCCCGCATGGTGCCTCGGGCTCCTTGGATGGAGCAAGAAGGCTCA GAGTACTGGGACCGGGAGACCCGATCTGCGCGCGATACAGCACA AATCTTTAGGGTCAACCTTCGAACATTGAGGGGCTACTACAACCAG AGTGAGGCAGGTTCCCATACGTTGCAATGGATGCATGGTTGCGAA CTTGGTCCAGATGGGAGGTTCCTCAGAGGTTATGAACAATTTGCTT ACGATGGAAAGGACTACCTTACACTCAATGAGGACCTCCGCAGCT GGACCGCCGTTGACACCGCTGCTCAAATCTCCGAGCAAAAGAGTA ATGACGCGTCAGAAGCGGAGCATCAGCGCGCCTACCTCGAGGAC ACGTGCGTTGAGTGGCTCCATAAATACCTGGAAAAAGGGAAAGAA ACCCTCCTCCACCTGGAGCCACCGAAGACCCACGTCACGCATCAC CCAATATCCGATCACGAGGCTACACTGAGGTGCTGGGCTCTCGGT HDRT27

[0254] 62 TTCTATCCGGCAGAAATAACCCTGACCTGGCAGCAGGATGGTGAG (B2 / B6) GGCCATACGCAGGACACGGAGCTTGTCGAGACTCGACCCGCCGG TGATGGAACATTCCAAAAATGGGCTGCTGTCGTTGTACCCTCTGGA GAAGAACAACGATACACCTGCCACGTGCAGCATGAGGGACTGCCG GAACCGGTAACCTTGCGCTGGAAGCCGGCGTCACAACCAACGATA CCCATCGTCGGCATTATAGCCGGACTTGTTCTCCTGGGAAGCGTG GTCAGCGGAGCTGTCGTAGCGGCCGTAATTTGGAGGAAAAAGTCT TCTGGGGGAAAGGGTGGGTCATACTCTAAAGCGGAGTGGTCTGAC TCAGCACAGGGTTCCGAATCCCACTCCCTGTGAGCGGCCGCGTCG AGTCTAGAGGGCCCGTTTAAACCCGCTGATCAGCCTCGACTGTGC CTTCTAGTTGCCAGCCATCTGTTGTTTGCCCCTCCCCCGTGCCTTC CTTGACCCTGGAAGGTGCCACTCCCACTGTCCTTTCCTAATAAAAT GAGGAAATTGCATCGCATTGTCTGAGTAGGTGTCATTCTATTCTGG GGGGTGGGGTGGGGCAGGACAGCAAGGGGGAGGATTGGGAAGA CAATAGCAGGCATGCTGGGGATGCGGTGGGCTCTATGGaaaaggtcta tggccatactaccctgaatgagtcccatcccatctgatataaacaatctgcatattgggattgtcagg gaatgttcttaaagatcagattagtggcacctgctgagatactgatgcacagcatggtttctgaacca gtagtttccctgcagttgagcagggagcagcagcagcacttgcacaaatacatatacactcttaac acttcttacctactggcttcctcTAGCTTTTGTGGCAGCTTCAGGTATATTTAGC ACTGAACGAACATCTCAAGAAGGTATAGGCCTTTGTTTGTAAGTCC

[0255] TGCTGTCCTAGCATCCTATAATCCTGGACTTCTCCAGTACTTTCTG GCTGGATTGGTATCTGAGGCTAGTAGGA _ aagctcatttggccagagtggaaatggaattgggagaaatcgatgaccaaatgtaaacacttggtg HDRT29

[0256] 63 cctgatatagcttgacaccaagttagccccaagtgaaataccctggcaatattaatgtgtcttttcccg (B2 / B9) atattcctcaggtactccaaagattcaggtttactcacgtcatccagcagagaatggaaagtcaaatt tcctgaattgctatgtgtctgggtttcatccatccgacattgaagttgacttactgaagaatggagaga gaattgaaaaagtggagcattcagacttgtctttcagcaaggactggtctttctatctcttgtactacact gaattcacccccactgaaaaagatgagtatgcAT GCCGT GT GAACCAT GT GACTT TGTCACAGCCTAAAATTGTGAAATGGGACAGGGACATGGGTGGGG GCGGATCAGGCGGCGGCGGTTCTGGGGGAGGCGGCTCAGGTGG CGGTGGGAGTGGATCCCATTCCCTGAAGTATTTCCACACCAGCGT TAGTCGGCCGGGAAGGGGAGAACCAAGATTCATTTCCGTCGGCTA TGTCGACGATACCCAATTTGTGCGATTTGATAATGACGCAGCTTCA CCCCGCATGGTGCCTCGGGCTCCTTGGATGGAGCAAGAAGGCTCA GAGTACTGGGACCGGGAGACCCGATCTGCGCGCGATACAGCACA AATCTTTAGGGTCAACCTTCGAACATTGAGGGGCTACTACAACCAG AGTGAGGCAGGTTCCCATACGTTGCAATGGATGCATGGTTGCGAA CTTGGTCCAGATGGGAGGTTCCTCAGAGGTTATGAACAATTTGCTT ACGATGGAAAGGACTACCTTACACTCAATGAGGACCTCCGCAGCT GGACCGCCGTTGACACCGCTGCTCAAATCTCCGAGCAAAAGAGTA ATGACGCGTCAGAAGCGGAGCATCAGCGCGCCTACCTCGAGGAC ACGTGCGTTGAGTGGCTCCATAAATACCTGGAAAAAGGGAAAGAA ACCCTCCTCCACCTGGAGCCACCGAAGACCCACGTCACGCATCAC CCAATATCCGATCACGAGGCTACACTGAGGTGCTGGGCTCTCGGT TTCTATCCGGCAGAAATAACCCTGACCTGGCAGCAGGATGGTGAG GGCCATACGCAGGACACGGAGCTTGTCGAGACTCGACCCGCCGG TGATGGAACATTCCAAAAATGGGCTGCTGTCGTTGTACCCTCTGGA GAAGAACAACGATACACCTGCCACGTGCAGCATGAGGGACTGCCG GAACCGGTAACCTTGCGCTGGAAGCCGGCGTCACAACCAACGATA CCCATCGTCGGCATTATAGCCGGACTTGTTCTCCTGGGAAGCGTG GTCAGCGGAGCTGTCGTAGCGGCCGTAATTTGGAGGAAAAAGTCT TCTGGGGGAAAGGGTGGGTCATACTCTAAAGCGGAGTGGTCTGAC TCAGCACAGGGTTCCGAATCCCACTCCCTGTGAGCGGCCGCGTCG AGTCTAGAGGGCCCGTTTAAACCCGCTGATCAGCCTCGACTGTGC CTTCTAGTTGCCAGCCATCTGTTGTTTGCCCCTCCCCCGTGCCTTC CTTGACCCTGGAAGGTGCCACTCCCACTGTCCTTTCCTAATAAAAT GAGGAAATTGCATCGCATTGTCTGAGTAGGTGTCATTCTATTCTGG GGGGTGGGGTGGGGCAGGACAGCAAGGGGGAGGATTGGGAAGA CAATAGCAGGCATGCTGGGGATGCGGTGGGCTCTATGGaaaaggtcta tggccatactaccctgaatgagtcccatcccatctgatataaacaatctgcatattgggattgtcagg gaatgttcttaaagatcagattagtggcacctgctgagatactgatgcacagcatggtttctgaacca gtagtttccctgcagttgagcagggagcagcagcagcacttgcacaaatacatatacactcttaac acttcttacctactggcttcctcTAGCTTTTGTGGCAGCTTCAGGTATATTTAGC ACTGAACGAACATCTCAAGAAGGTATAGGCCTTTGTTTGTAAGTCC TGCTGTCCTAGCATCCTATAATCCTGGACTTCTCCAGTACTTTCTG GCTGGATTGGTATCTGAGGCTAGTAGGA _ aagctcatttggccagagtggaaatggaattgggagaaatcgatgaccaaatgtaaacacttggtg cctgatatagcttgacaccaagttagccccaagtgaaataccctggcaatattaatgtgtcttttcccg atattcctcaggtactccaaagattcaggtttactcacgtcatccagcagagaatggaaagtcaaatt tcctgaattgctatgtgtctgggtttcatccatccgacattgaagttgacttactgaagaatggagaga gaattgaaaaagtggagcattcagacttgtctttcagcaaggactggtctttctatctcttgtactacact gaattcacccccactgaaaaagatgagtatgctTGCAGGGTTAACCATGTCACATT GTCACAACCTAAAATTGTGAAATGGGACAGGGACATGGGTGGGGG CGGATCAGGCGGCGGCGGTTCTGGGGGAGGCGGCTCAGGTGGC HDRT26

[0257] 64 GGTGGGAGTGGATCCCATTCCCTGAAGTATTTCCACACCAGCGTTA (B4 / B8) GTCGGCCGGGAAGGGGAGAACCAAGATTCATTTCCGTCGGCTATG TCGACGATACCCAATTTGTGCGATTTGATAATGACGCAGCTTCACC CCGCATGGTGCCTCGGGCTCCTTGGATGGAGCAAGAAGGCTCAGA GTACTGGGACCGGGAGACCCGATCTGCGCGCGATACAGCACAAAT CTTTAGGGTCAACCTTCGAACATTGAGGGGCTACTACAACCAGAGT GAGGCAGGTTCCCATACGTTGCAATGGATGCATGGTTGCGAACTT GGTCCAGATGGGAGGTTCCTCAGAGGTTATGAACAATTTGCTTACG ATGGAAAGGACTACCTTACACTCAATGAGGACCTCCGCAGCTGGA CCGCCGTTGACACCGCTGCTCAAATCTCCGAGCAAAAGAGTAATG ACGCGTCAGAAGCGGAGCATCAGCGCGCCTACCTCGAGGACACG TGCGTTGAGTGGCTCCATAAATACCTGGAAAAAGGGAAAGAAACC CTCCTCCACCTGGAGCCACCGAAGACCCACGTCACGCATCACCCA ATATCCGATCACGAGGCTACACTGAGGTGCTGGGCTCTCGGTTTC TATCCGGCAGAAATAACCCTGACCTGGCAGCAGGATGGTGAGGGC CATACGCAGGACACGGAGCTTGTCGAGACTCGACCCGCCGGTGAT GGAACATTCCAAAAATGGGCTGCTGTCGTTGTACCCTCTGGAGAA GAACAACGATACACCTGCCACGTGCAGCATGAGGGACTGCCGGAA CCGGTAACCTTGCGCTGGAAGCCGGCGTCACAACCAACGATACCC ATCGTCGGCATTATAGCCGGACTTGTTCTCCTGGGAAGCGTGGTC AGCGGAGCTGTCGTAGCGGCCGTAATTTGGAGGAAAAAGTCTTCT GGGGGAAAGGGTGGGTCATACTCTAAAGCGGAGTGGTCTGACTCA GCACAGGGTTCCGAATCCCACTCCCTGTGAGCGGCCGCGTCGAGT CTAGAGGGCCCGTTTAAACCCGCTGATCAGCCTCGACTGTGCCTT CTAGTTGCCAGCCATCTGTTGTTTGCCCCTCCCCCGTGCCTTCCTT GACCCTGGAAGGTGCCACTCCCACTGTCCTTTCCTAATAAAATGAG GAAATTGCATCGCATTGTCTGAGTAGGTGTCATTCTATTCTGGGGG GTGGGGTGGGGCAGGACAGCAAGGGGGAGGATTGGGAAGACAAT AGCAGGCAT GCT GGGGAT GCGGT GGGCT CT AT GGccaagatagttaagt ggggtaagtcttacattcttttgtaagctgctgaaagttgtgtatgagtagtcatatcataaagctgctttg atataaaaaaggtctatggccatactaccctgaatgagtcccatcccatctgatataaacaatctgc atattgggattgtcagggaatgttcttaaagatcagattagtggcacctgctgagatactgatgcaca gcatggtttctgaaccagtagtttccctgcagttgagcagggagcagcagcagcacttgcacaaat acatatacactcttaacacttcttacctactggcttcctctagcttttgtggcagcttcaggta _ aagctcatttggccagagtggaaatggaattgggagaaatcgatgaccaaatgtaaacacttggtg cctgatatagcttgacaccaagttagccccaagtgaaataccctggcaatattaatgtgtcttttcccg atattcctcaggtactccaaagattcaggtttactcacgtcatccagcagagaatggaaagtcaaatt tcctgaattgctatgtgtctgggtttcatccatccgacattgaagttgacttactgaagaatggagaga gaattgaaaaagtggagcattcagacttgtctttcagcaaggactggtctttctatctcttgtactacact gaattcacccccactgaaaaagatgagtatgctTGCAGGGTTAACCATGTCACATT GTCACAACCTAAAATTGTGAAATGGGACAGGGACATGGGTGGGGG CGGATCAGGCGGCGGCGGTTCTGGGGGAGGCGGCTCAGGTGGC GGTGGGAGTGGATCCCATTCCCTGAAGTATTTCCACACCAGCGTTA GTCGGCCGGGAAGGGGAGAACCAAGATTCATTTCCGTCGGCTATG TCGACGATACCCAATTTGTGCGATTTGATAATGACGCAGCTTCACC CCGCATGGTGCCTCGGGCTCCTTGGATGGAGCAAGAAGGCTCAGA GTACTGGGACCGGGAGACCCGATCTGCGCGCGATACAGCACAAAT CTTTAGGGTCAACCTTCGAACATTGAGGGGCTACTACAACCAGAGT GAGGCAGGTTCCCATACGTTGCAATGGATGCATGGTTGCGAACTT GGTCCAGATGGGAGGTTCCTCAGAGGTTATGAACAATTTGCTTACG HDRT26

[0258] 65 ATGGAAAGGACTACCTTACACTCAATGAGGACCTCCGCAGCTGGA (B5 / B8) CCGCCGTTGACACCGCTGCTCAAATCTCCGAGCAAAAGAGTAATG ACGCGTCAGAAGCGGAGCATCAGCGCGCCTACCTCGAGGACACG TGCGTTGAGTGGCTCCATAAATACCTGGAAAAAGGGAAAGAAACC

[0259] CTCCTCCACCTGGAGCCACCGAAGACCCACGTCACGCATCACCCA ATATCCGATCACGAGGCTACACTGAGGTGCTGGGCTCTCGGTTTC TATCCGGCAGAAATAACCCTGACCTGGCAGCAGGATGGTGAGGGC CATACGCAGGACACGGAGCTTGTCGAGACTCGACCCGCCGGTGAT GGAACATTCCAAAAATGGGCTGCTGTCGTTGTACCCTCTGGAGAA GAACAACGATACACCTGCCACGTGCAGCATGAGGGACTGCCGGAA CCGGTAACCTTGCGCTGGAAGCCGGCGTCACAACCAACGATACCC ATCGTCGGCATTATAGCCGGACTTGTTCTCCTGGGAAGCGTGGTC AGCGGAGCTGTCGTAGCGGCCGTAATTTGGAGGAAAAAGTCTTCT GGGGGAAAGGGTGGGTCATACTCTAAAGCGGAGTGGTCTGACTCA GCACAGGGTTCCGAATCCCACTCCCTGTGAGCGGCCGCGTCGAGT CTAGAGGGCCCGTTTAAACCCGCTGATCAGCCTCGACTGTGCCTT CTAGTTGCCAGCCATCTGTTGTTTGCCCCTCCCCCGTGCCTTCCTT GACCCTGGAAGGTGCCACTCCCACTGTCCTTTCCTAATAAAATGAG GAAATTGCATCGCATTGTCTGAGTAGGTGTCATTCTATTCTGGGGG GTGGGGTGGGGCAGGACAGCAAGGGGGAGGATTGGGAAGACAAT AGCAGGCAT GCT GGGGAT GCGGT GGGCT CT AT GGccaagatagttaagt ggggtaagtcttacattcttttgtaagctgctgaaagttgtgtatgagtagtcatatcataaagctgctttg atataaaaaaggtctatggccatactaccctgaatgagtcccatcccatctgatataaacaatctgc atattgggattgtcagggaatgttcttaaagatcagattagtggcacctgctgagatactgatgcaca gcatggtttctgaaccagtagtttccctgcagttgagcagggagcagcagcagcacttgcacaaat acatatacactcttaacacttcttacctactggcttcctctagcttttgtggcagcttcaggta

[0260] TRAC

[0261] 66 AAAGTCAGATTTGTTGCTCC sense T3.1 TRAC

[0262] 67 CTTACCTGGGCTGGGGAAGA sense T4.1 TRAC

[0263] 68 AGCTGCCCTTACCTGGGCTG sense T5.1 TRAC

[0264] 69 CACCAAAGCTGCCCTTACCT sense T6.1 CD3E

[0265] 70 TGTCCTTCCCTCTTGCTGCT sense E1 CD3E

[0266] 71 TGCATTCCTAAACAGAAAAA sense E2 CD3E

[0267] 72 AAAGTGTTAGTAAATGTTGT sense E3 CD3E

[0268] 73 GTTGGCGTTTGGGGGCAAGA antisense E4 CD3E

[0269] 74 TCATTCTTTGTAGTTATGAA antisense E5 B2M B2

[0270] 75 CTTGCCCCACTTAACTATCT retargetin g sgRNA B2M (B2)

[0271] 76 CTTACCCCACTTAACTATCT Nuclease gRNA CIITA

[0272] 77 CACTCACCTTAGCCTGAGCA nuclease gRNA HLA-A

[0273] 78 CGAGCCAGAGGATGGAGCCG

[0274] CBE #1

[0275] HLA-A

[0276] 79 TACCACCAGTACGCCTACGA

[0277] CBE #2

[0278] HLA-A

[0279] 80 ACCCGCCCAGGTCTGGGTCA

[0280] CBE #2

[0281] HLA-A

[0282] 81 TGCGGAGCCACTCCACGCAC

[0283] CBE #4 HLA-B

[0284] 82 ACTCTCAGGCTGCGTGTAAG

[0285] CBE #1

[0286] HLA-B

[0287] 83 GACCTGGCAGCGGGATGGCG

[0288] CBE #2

[0289] HLA-B

[0290] 84 ACCGGCCCAGGTCTCGGTCA

[0291] CBE #3

[0292] HLA-B

[0293] 85 ACCGGCCCAGGTCTCGGTCA

[0294] CBE #4

[0295] HLA-C

[0296] 86 GATCACCCAGCGCAAGTGGG

[0297] CBE #1

[0298] HLA-C

[0299] 87 TATGACCAGTCCGCCTACGA

[0300] CBE #2

[0301] HLA-C

[0302] 88 ACAGGCCCAGGTCTCGGTC

[0303] CBE #3 cctattaaataaaagaataagcagtattattaagtagccctgcatttcaggtttccttgagtggcaggc caggcctggccgtgaacgttcactgaaatcatggcctcttggccaagattgatagcttgtgcctgtcc ctgagtcccagtccatcacgagcagctggtttctaagatgctatttcccgtataaagcatgagaccgt gacttgccagccccacagagccccgcccttgtccatcactggcatctggactccagcctgggttgg ggcaaagagggaaatgagatcatgtcctaaccctgatcctcttgtcccacagatatccagaaccct gaccctgccgtgtaccagctgagagactctaaatccagtgacaagtctgtctgcctattcaccGCC ACCAACTTCTCTCTTTTGAAGCAGGCCGGAGATGTGGAAGAAAATC

[0304] CTGGGCCCATGGCTCTTCCTGTGACTGCCCTTCTGCTGCCCCTGG

[0305] CTCTCCTGCTCCATGCTGCCCGGCCAGACATCCAGATGACACAGA

[0306] CAACCAGCAGCCTCAGCGCCAGCCTGGGGGACAGAGTGACCATTA

[0307] GCTGCCGGGCCTCTCAGGACATCAGCAAATACCTGAACTGGTACC

[0308] AGCAGAAACCAGATGGCACTGTCAAGCTGCTGATTTACCACACCTC

[0309] CAGGCTCCACAGCGGCGTGCCCAGTCGCTTCAGCGGCAGTGGGA

[0310] 89 HDTR10

[0311] GCGGGACAGATTATTCCCTCACAATCTCCAACCTGGAGCAGGAAG

[0312] ATATTGCCACATACTTCTGCCAGCAAGGCAACACCCTGCCATACAC

[0313] ATTTGGAGGCGGCACCAAATTGGAGATCACCGGCGGTGGTGGATC

[0314] TGGAGGAGGGGGCAGCGGAGGTGGCGGCTCTGAGGTGAAACTGC

[0315] AGGAGAGTGGCCCTGGCCTGGTGGCTCCCAGCCAGAGCCTTTCT

[0316] GTCACCTGCACCGTGTCTGGGGTGTCCCTGCCTGACTATGGAGTC

[0317] TCCTGGATCCGGCAGCCTCCAAGAAAAGGACTGGAATGGCTGGGC

[0318] GTCATCTGGGGAAGTGAGACCACCTACTATAATTCAGCCCTCAAGT

[0319] CCCGGCTCACCATCATTAAGGACAACTCCAAATCCCAGGTGTTCCT

[0320] GAAGATGAATTCTCTCCAGACTGATGACACAGCCATCTACTACTGT

[0321] GCCAAGCATTATTATTACGGCGGGTCCTATGCCATGGACTACTGG

[0322] GGCCAGGGGACCAGTGTCACTGTTTCTTCTGAAAGCAAGTATGGC

[0323] CCCCCATGCCCTCCCTGCCCCGGACAGCCGCGGGAGCCCCAGGT CTACACCCTCCCACCCTCCAGAGATGAGCTGACCAAGAACCAGGT

[0324] TTCACTGACATGCCTGGTGAAGGGCTTCTACCCCTCTGACATTGCT

[0325] GTGGAGTGGGAGAGCAATGGGCAGCCAGAGAACAACTACAAGAC

[0326] CACACCTCCTGTCCTGGACAGCGACGGCTCCTTCTTCCTCTACAGT

[0327] AAGCTCACTGTGGACAAGAGCCGCTGGCAGCAGGGCAATGTCTTC

[0328] TCCTGCAGCGTGATGCACGAGGCCCTGCACAATGCCTATACCCAG

[0329] AAGTCACTCTCTCTGAGCCCTGGAAAGAAGGATCCCAAGTTTTGGG

[0330] TCCTCGTGGTGGTCGGGGGTGTGCTGGCCTGCTACAGCTTGCTGG

[0331] TCACAGTGGCCTTCATCATCTTCTGGGTGCGCTCCAAGAGGAGCC

[0332] GGCTGCTTCACAGTGATTACATGAACATGACCCCCAGGAGGCCAG

[0333] GACCCACCAGGAAGCACTACCAGCCCTACGCTCCCCCGCGGGAC

[0334] TTTGCTGCTTACCGCAGCAGGGTCAAATTTTCTAGATCTGCAGATG

[0335] CGCCGGCCTATCAAcaaggccagaaccagctcTATAACGAGCTCAATCTA

[0336] GGACGAAGAGAGGAGTACGATGTTTTGGACAAGAGACGTGGCCG

[0337] GGACCCTGAGATGGGGGGAAAGCCGAGAAGGAAGAACCCTCAGG

[0338] AAGGCCTGTACAATGAACTGCAGAAAGATAAGATGGCGGAGGCCT

[0339] ACAGTGAGATTGGGATGAAAGGCGAGCGCCGGAGGGGCAAGGGG

[0340] CACGATGGCCTTTACCAGGGTCTCAGTACAGCCACCAAGGACACC

[0341] TACGACGCCCTTCACATGCAGGCCCTGCCCCCTCGCTAAtaagaattct aactagagctcgctgatcagcctcgactgtgccttctagttgccagccatctgttgtttgcccctccccc gtgccttccttgaccctggaaggtgccactcccactgtcctttcctaataaaatgaggaaattgcatcg cattgtctgagtaggtgtcattctattctggggggtggggtggggcaggacagcaagggggaggat tgggaagagaatagcaggcatgctggggactttggtgccttcgcaggctgtttccttgcttcaggaat ggccaggttctgcccagagctctggtcaatgatgtctaaaactcctctgattggtggtctcggccttatc cattgccaccaaaaccctctttttactaagaaacagtgagccttgttctggcagtccagagaatgac acgggaaaaaagcagatgaagagaaggtggcaggagagggcacgtggcccagcctcagtct ctccaactgagttcctgcctgcctgcctttgctcagactgtttgccccttactgctcttctaggcctcattct aagccccttctccaagttgcctctccttatttctccctgtctgccaaaaaatctttcccagctcactaagt cagtctcacgcagtcactc cctattaaataaaagaataagcagtattattaagtagccctgcatttcaggtttccttgagtggcaggc caggcctggccgtgaacgttcactgaaatcatggcctcttggccaagattgatagcttgtgcctgtcc ctgagtcccagtccatcacgagcagctggtttctaagatgctatttcccgtataaagcatgagaccgt gacttgccagccccacagagccccgcccttgtccatcactggcatctggactccagcctgggttgg ggcaaagagggaaatgagatcatgtcctaaccctgatcctcttgtcccacagatatccagaaccct gaccctgccgtgtaccagctgagagactctaaatccagtgacaagtctgtctgcctattcaccgatttt gattctcaaacaaatgtgtcacaaagtGGAT CCGGAGCCACCAACTT CAGCCT G

[0342] 90 CTGAAGCAGGCCGGCGACGTGGAAGAAAATCCTGGGCCCATGGC HDTR11

[0343] TCTTCCTGTGACTGCCCTTCTGCTGCCCCTGGCTCTCCTGCTCCAT

[0344] GCTGCCCGGCCAGACATCCAGATGACACAGACAACCAGCAGCCTC

[0345] AGCGCCAGCCTGGGGGACAGAGTGACCATTAGCTGCCGGGCCTC

[0346] TCAGGACATCAGCAAATACCTGAACTGGTACCAGCAGAAACCAGAT

[0347] GGCACTGTCAAGCTGCTGATTTACCACACCTCCAGGCTCCACAGC

[0348] GGCGTGCCCAGTCGCTTCAGCGGCAGTGGGAGCGGGACAGATTA

[0349] TTCCCTCACAATCTCCAACCTGGAGCAGGAAGATATTGCCACATAC TTCTGCCAGCAAGGCAACACCCTGCCATACACATTTGGAGGCGGC

[0350] ACCAAATTGGAGATCACCGGCGGTGGTGGATCTGGAGGAGGGGG

[0351] CAGCGGAGGTGGCGGCTCTGAGGTGAAACTGCAGGAGAGTGGCC

[0352] CTGGCCTGGTGGCTCCCAGCCAGAGCCTTTCTGTCACCTGCACCG

[0353] TGTCTGGGGTGTCCCTGCCTGACTATGGAGTCTCCTGGATCCGGC

[0354] AGCCTCCAAGAAAAGGACTGGAATGGCTGGGCGTCATCTGGGGAA

[0355] GTGAGACCACCTACTATAATTCAGCCCTCAAGTCCCGGCTCACCAT

[0356] CATTAAGGACAACTCCAAATCCCAGGTGTTCCTGAAGATGAATTCT

[0357] CTCCAGACTGATGACACAGCCATCTACTACTGTGCCAAGCATTATT

[0358] ATTACGGCGGGTCCTATGCCATGGACTACTGGGGCCAGGGGACCA

[0359] GTGTCACTGTTTCTTCTGAAAGCAAGTATGGCCCCCCATGCCCTCC

[0360] CTGCCCCGGACAGCCGCGGGAGCCCCAGGTCTACACCCTCCCAC

[0361] CCTCCAGAGATGAGCTGACCAAGAACCAGGTTTCACTGACATGCC

[0362] TGGTGAAGGGCTTCTACCCCTCTGACATTGCTGTGGAGTGGGAGA

[0363] GCAATGGGCAGCCAGAGAACAACTACAAGACCACACCTCCTGTCC

[0364] TGGACAGCGACGGCTCCTTCTTCCTCTACAGTAAGCTCACTGTGGA

[0365] CAAGAGCCGCTGGCAGCAGGGCAATGTCTTCTCCTGCAGCGTGAT

[0366] GCACGAGGCCCTGCACAATGCCTATACCCAGAAGTCACTCTCTCT

[0367] GAGCCCTGGAAAGAAGGATCCCAAGTTTTGGGTCCTCGTGGTGGT

[0368] CGGGGGTGTGCTGGCCTGCTACAGCTTGCTGGTCACAGTGGCCTT

[0369] CATCATCTTCTGGGTGCGCTCCAAGAGGAGCCGGCTGCTTCACAG

[0370] TGATTACATGAACATGACCCCCAGGAGGCCAGGACCCACCAGGAA

[0371] GCACTACCAGCCCTACGCTCCCCCGCGGGACTTTGCTGCTTACCG

[0372] CAGCAGGGTCAAATTTTCTAGATCTGCAGATGCGCCGGCCTATCAA caaggccagaaccagctcTATAACGAGCTCAATCTAGGACGAAGAGAGGA

[0373] GTACGATGTTTTGGACAAGAGACGTGGCCGGGACCCTGAGATGGG

[0374] GGGAAAGCCGAGAAGGAAGAACCCTCAGGAAGGCCTGTACAATGA

[0375] ACTGCAGAAAGATAAGATGGCGGAGGCCTACAGTGAGATTGGGAT

[0376] GAAAGGCGAGCGCCGGAGGGGCAAGGGGCACGATGGCCTTTACC

[0377] AGGGTCTCAGTACAGCCACCAAGGACACCTACGACGCCCTTCACA

[0378] T GCAGGCCCT GCCCCCT CGCT AAtaagaattctaactagagctcgctgatcagcct cgactgtgccttctagttgccagccatctgttgtttgcccctcccccgtgccttccttgaccctggaaggt gccactcccactgtcctttcctaataaaatgaggaaattgcatcgcattgtctgagtaggtgtcattcta ttctggggggtggggtggggcaggacagcaagggggaggattgggaagagaatagcaggcat gctggggagcaacaaatctgactttgcatgtgcaaacgccttcaacaacagcattattccagaaga caccttcttccccagcccaggtaagggcagctttggtgccttcgcaggctgtttccttgcttcaggaat ggccaggttctgcccagagctctggtcaatgatgtctaaaactcctctgattggtggtctcggccttatc cattgccaccaaaaccctctttttactaagaaacagtgagccttgttctggcagtccagagaatgac acgggaaaaaagcagatgaagagaaggtggcaggagagggcacgtggcccagcctcagtct ctccaactgagttcctgcctgcctgcctttgctcagactgtttgccccttactgctcttctaggcctc tgatagcttgtgcctgtccctgagtcccagtccatcacgagcagctggtttctaagatgctatttcccgt ataaagcatgagaccgtgacttgccagccccacagagccccgcccttgtccatcactggcatctgg

[0379] 91 HDTR12 actccagcctgggttggggcaaagagggaaatgagatcatgtcctaaccctgatcctcttgtcccac agatatccagaaccctgaccctgccgtgtaccagctgagagactctaaatccagtgacaagtctgt ctgcctattcaccgattttgattctcaaacaaatgtgtcacaaagtaaggattctgatgtgtatatcaca gacaaaactgtgctagacatgaggtctGGATCCGGAGCCACCAACTTCAGCCT

[0380] GCTGAAGCAGGCCGGCGACGTGGAAGAAAATCCTGGGCCCATGG

[0381] CTCTTCCTGTGACTGCCCTTCTGCTGCCCCTGGCTCTCCTGCTCCA

[0382] TGCTGCCCGGCCAGACATCCAGATGACACAGACAACCAGCAGCCT

[0383] CAGCGCCAGCCTGGGGGACAGAGTGACCATTAGCTGCCGGGCCT

[0384] CTCAGGACATCAGCAAATACCTGAACTGGTACCAGCAGAAACCAG

[0385] ATGGCACTGTCAAGCTGCTGATTTACCACACCTCCAGGCTCCACAG

[0386] CGGCGTGCCCAGTCGCTTCAGCGGCAGTGGGAGCGGGACAGATT

[0387] ATTCCCTCACAATCTCCAACCTGGAGCAGGAAGATATTGCCACATA

[0388] CTTCTGCCAGCAAGGCAACACCCTGCCATACACATTTGGAGGCGG

[0389] CACCAAATTGGAGATCACCGGCGGTGGTGGATCTGGAGGAGGGG

[0390] GCAGCGGAGGTGGCGGCTCTGAGGTGAAACTGCAGGAGAGTGGC

[0391] CCTGGCCTGGTGGCTCCCAGCCAGAGCCTTTCTGTCACCTGCACC

[0392] GTGTCTGGGGTGTCCCTGCCTGACTATGGAGTCTCCTGGATCCGG

[0393] CAGCCTCCAAGAAAAGGACTGGAATGGCTGGGCGTCATCTGGGGA

[0394] AGTGAGACCACCTACTATAATTCAGCCCTCAAGTCCCGGCTCACCA

[0395] TCATTAAGGACAACTCCAAATCCCAGGTGTTCCTGAAGATGAATTC

[0396] TCTCCAGACTGATGACACAGCCATCTACTACTGTGCCAAGCATTAT

[0397] TATTACGGCGGGTCCTATGCCATGGACTACTGGGGCCAGGGGACC

[0398] AGTGTCACTGTTTCTTCTGAAAGCAAGTATGGCCCCCCATGCCCTC

[0399] CCTGCCCCGGACAGCCGCGGGAGCCCCAGGTCTACACCCTCCCA

[0400] CCCTCCAGAGATGAGCTGACCAAGAACCAGGTTTCACTGACATGC

[0401] CTGGTGAAGGGCTTCTACCCCTCTGACATTGCTGTGGAGTGGGAG

[0402] AGCAATGGGCAGCCAGAGAACAACTACAAGACCACACCTCCTGTC

[0403] CTGGACAGCGACGGCTCCTTCTTCCTCTACAGTAAGCTCACTGTG

[0404] GACAAGAGCCGCTGGCAGCAGGGCAATGTCTTCTCCTGCAGCGTG

[0405] ATGCACGAGGCCCTGCACAATGCCTATACCCAGAAGTCACTCTCTC

[0406] TGAGCCCTGGAAAGAAGGATCCCAAGTTTTGGGTCCTCGTGGTGG

[0407] TCGGGGGTGTGCTGGCCTGCTACAGCTTGCTGGTCACAGTGGCCT

[0408] TCATCATCTTCTGGGTGCGCTCCAAGAGGAGCCGGCTGCTTCACA

[0409] GTGATTACATGAACATGACCCCCAGGAGGCCAGGACCCACCAGGA

[0410] AGCACTACCAGCCCTACGCTCCCCCGCGGGACTTTGCTGCTTACC

[0411] GCAGCAGGGTCAAATTTTCTAGATCTGCAGATGCGCCGGCCTATC

[0412] AAcaaggccagaaccagctcTATAACGAGCTCAATCTAGGACGAAGAGAG

[0413] GAGTACGATGTTTTGGACAAGAGACGTGGCCGGGACCCTGAGATG

[0414] GGGGGAAAGCCGAGAAGGAAGAACCCTCAGGAAGGCCTGTACAA

[0415] TGAACTGCAGAAAGATAAGATGGCGGAGGCCTACAGTGAGATTGG

[0416] GATGAAAGGCGAGCGCCGGAGGGGCAAGGGGCACGATGGCCTTT

[0417] ACCAGGGTCTCAGTACAGCCACCAAGGACACCTACGACGCCCTTC

[0418] ACAT GCAGGCCCT GCCCCCT CGCT AAtaagaattctaactagagctcgctgatca gcctcgactgtgccttctagttgccagccatctgttgtttgcccctcccccgtgccttccttgaccctgga aggtgccactcccactgtcctttcctaataaaatgaggaaattgcatcgcattgtctgagtaggtgtca ttctattctggggggtggggtggggcaggacagcaagggggaggattgggaagagaatagcag gcatgctggggagcaacaaatctgactttgcatgtgcaaacgccttcaacaacagcattattccag aagacaccttcttccccagcccaggtaagggcagctttggtgccttcgcaggctgtttccttgcttcag gaatggccaggttctgcccagagctctggtcaatgatgtctaaaactcctctgattggtggtctcggc cttatccattgccaccaaaaccctctttttactaagaaacagtgagccttgttctggcagtccagagaa tgacacgggaaaaaagcagatgaagagaaggtggcaggagagggcacgtggcccagcctca gtctctccaactgagttcctgcctgcctgcctttgctcagactgtttgccccttactgctcttctaggcctc tgatagcttgtgcctgtccctgagtcccagtccatcacgagcagctggtttctaagatgctatttcccgt ataaagcatgagaccgtgacttgccagccccacagagccccgcccttgtccatcactggcatctgg actccagcctgggttggggcaaagagggaaatgagatcatgtcctaaccctgatcctcttgtcccac agatatccagaaccctgaccctgccgtgtaccagctgagagactctaaatccagtgacaagtctgt ctgcctattcaccgattttgattctcaaacaaatgtgtcacaaagtaaggattctgatgtgtatatcaca gacaaaactgtgctagacatgaggtctatggacttcaagagcaacagtgctgtgGGATCCGG AGCCACCAACTTCAGCCTGCTGAAGCAGGCCGGCGACGTGGAAG

[0419] AAAATCCTGGGCCCATGGCTCTTCCTGTGACTGCCCTTCTGCTGCC

[0420] CCTGGCTCTCCTGCTCCATGCTGCCCGGCCAGACATCCAGATGAC

[0421] ACAGACAACCAGCAGCCTCAGCGCCAGCCTGGGGGACAGAGTGA

[0422] CCATTAGCTGCCGGGCCTCTCAGGACATCAGCAAATACCTGAACT

[0423] GGTACCAGCAGAAACCAGATGGCACTGTCAAGCTGCTGATTTACC

[0424] ACACCTCCAGGCTCCACAGCGGCGTGCCCAGTCGCTTCAGCGGC

[0425] AGTGGGAGCGGGACAGATTATTCCCTCACAATCTCCAACCTGGAG

[0426] CAGGAAGATATTGCCACATACTTCTGCCAGCAAGGCAACACCCTG

[0427] CCATACACATTTGGAGGCGGCACCAAATTGGAGATCACCGGCGGT

[0428] GGTGGATCTGGAGGAGGGGGCAGCGGAGGTGGCGGCTCTGAGG

[0429] TGAAACTGCAGGAGAGTGGCCCTGGCCTGGTGGCTCCCAGCCAG

[0430] AGCCTTTCTGTCACCTGCACCGTGTCTGGGGTGTCCCTGCCTGAC

[0431] 92 HDRT18

[0432] TATGGAGTCTCCTGGATCCGGCAGCCTCCAAGAAAAGGACTGGAA

[0433] TGGCTGGGCGTCATCTGGGGAAGTGAGACCACCTACTATAATTCA

[0434] GCCCTCAAGTCCCGGCTCACCATCATTAAGGACAACTCCAAATCCC

[0435] AGGTGTTCCTGAAGATGAATTCTCTCCAGACTGATGACACAGCCAT

[0436] CTACTACTGTGCCAAGCATTATTATTACGGCGGGTCCTATGCCATG

[0437] GACTACTGGGGCCAGGGGACCAGTGTCACTGTTTCTTCTGAAAGC

[0438] AAGTATGGCCCCCCATGCCCTCCCTGCCCCGGACAGCCGCGGGA

[0439] GCCCCAGGTCTACACCCTCCCACCCTCCAGAGATGAGCTGACCAA

[0440] GAACCAGGTTTCACTGACATGCCTGGTGAAGGGCTTCTACCCCTCT

[0441] GACATTGCTGTGGAGTGGGAGAGCAATGGGCAGCCAGAGAACAAC

[0442] TACAAGACCACACCTCCTGTCCTGGACAGCGACGGCTCCTTCTTC

[0443] CTCTACAGTAAGCTCACTGTGGACAAGAGCCGCTGGCAGCAGGGC

[0444] AATGTCTTCTCCTGCAGCGTGATGCACGAGGCCCTGCACAATGCC

[0445] TATACCCAGAAGTCACTCTCTCTGAGCCCTGGAAAGAAGGATCCCA

[0446] AGTTTTGGGTCCTCGTGGTGGTCGGGGGTGTGCTGGCCTGCTACA

[0447] GCTTGCTGGTCACAGTGGCCTTCATCATCTTCTGGGTGCGCTCCAA

[0448] GAGGAGCCGGCTGCTTCACAGTGATTACATGAACATGACCCCCAG

[0449] GAGGCCAGGACCCACCAGGAAGCACTACCAGCCCTACGCTCCCC

[0450] CGCGGGACTTTGCTGCTTACCGCAGCAGGGTCAAATTTTCTAGATC TGCAGATGCGCCGGCCTATCAAcaaggccagaaccagctcTATAACGAGC

[0451] TCAATCTAGGACGAAGAGAGGAGTACGATGTTTTGGACAAGAGAC

[0452] GTGGCCGGGACCCTGAGATGGGGGGAAAGCCGAGAAGGAAGAAC

[0453] CCTCAGGAAGGCCTGTACAATGAACTGCAGAAAGATAAGATGGCG

[0454] GAGGCCTACAGTGAGATTGGGATGAAAGGCGAGCGCCGGAGGGG

[0455] CAAGGGGCACGATGGCCTTTACCAGGGTCTCAGTACAGCCACCAA

[0456] GGACACCTACGACGCCCTTCACATGCAGGCCCTGCCCCCTCGCTA

[0457] Ataagaattctaactagagctcgctgatcagcctcgactgtgccttctagttgccagccatctgttgttt gcccctcccccgtgccttccttgaccctggaaggtgccactcccactgtcctttcctaataaaatgag gaaattgcatcgcattgtctgagtaggtgtcattctattctggggggtggggtggggcaggacagca agggggaggattgggaagagaatagcaggcatgctggggagcaacaaatctgactttgcatgtg caaacgccttcaacaacagcattattccagaagacaccttcttccccagcccaggtaagggcagc tttggtgccttcgcaggctgtttccttgcttcaggaatggccaggttctgcccagagctctggtcaatga tgtctaaaactcctctgattggtggtctcggccttatccattgccaccaaaaccctctttttactaagaaa cagtgagccttgttctggcagtccagagaatgacacgggaaaaaagcagatgaagagaaggtg gcaggagagggcacgtggcccagcctcagtctctccaactgagttcctgcctgcctgcctttgctca gactgtttgccccttactgctcttctaggcctc cctattaaataaaagaataagcagtattattaagtagccctgcatttcaggtttccttgagtggcaggc caggcctggccgtgaacgttcactgaaatcatggcctcttggccaagattgatagcttgtgcctgtcc ctgagtcccagtccatcacgagcagctggtttctaagatgctatttcccgtataaagcatgagaccgt gacttgccagccccacagagccccgcccttgtccatcactggcatctggactccagcctgggttgg ggcaaagagggaaatgagatcatgtcctaaccctgatcctcttgtcccacagatatccagaaccct gaccctgccgtgtaccagctgagagactctaaatccagtgacaagtctgtctgcctattcaccgatttt gattctcaaacaaatgtgtcacaaagtGGAT CCGGAGCCACCAACTT CAGCCT G CTGAAGCAGGCCGGCGACGTGGAAGAAAATCCTGGGCCCATGGC

[0458] TCTTCCTGTGACTGCCCTTCTGCTGCCCCTGGCTCTCCTGCTCCAT

[0459] GCTGCCCGGCCAGACATCCAGATGACACAGACAACCAGCAGCCTC

[0460] AGCGCCAGCCTGGGGGACAGAGTGACCATTAGCTGCCGGGCCTC

[0461] TCAGGACATCAGCAAATACCTGAACTGGTACCAGCAGAAACCAGAT

[0462] GGCACTGTCAAGCTGCTGATTTACCACACCTCCAGGCTCCACAGC

[0463] 93 GGCGTGCCCAGTCGCTTCAGCGGCAGTGGGAGCGGGACAGATTA HDTR13

[0464] TTCCCTCACAATCTCCAACCTGGAGCAGGAAGATATTGCCACATAC

[0465] TTCTGCCAGCAAGGCAACACCCTGCCATACACATTTGGAGGCGGC

[0466] ACCAAATTGGAGATCACCGGCGGTGGTGGATCTGGAGGAGGGGG

[0467] CAGCGGAGGTGGCGGCTCTGAGGTGAAACTGCAGGAGAGTGGCC

[0468] CTGGCCTGGTGGCTCCCAGCCAGAGCCTTTCTGTCACCTGCACCG

[0469] TGTCTGGGGTGTCCCTGCCTGACTATGGAGTCTCCTGGATCCGGC

[0470] AGCCTCCAAGAAAAGGACTGGAATGGCTGGGCGTCATCTGGGGAA

[0471] GTGAGACCACCTACTATAATTCAGCCCTCAAGTCCCGGCTCACCAT

[0472] CATTAAGGACAACTCCAAATCCCAGGTGTTCCTGAAGATGAATTCT

[0473] CTCCAGACTGATGACACAGCCATCTACTACTGTGCCAAGCATTATT

[0474] ATTACGGCGGGTCCTATGCCATGGACTACTGGGGCCAGGGGACCA

[0475] GTGTCACTGTTTCTTCTGAAAGCAAGTATGGCCCCCCATGCCCTCC

[0476] CTGCCCCGGACAGCCGCGGGAGCCCCAGGTCTACACCCTCCCAC CCTCCAGAGATGAGCTGACCAAGAACCAGGTTTCACTGACATGCC

[0477] TGGTGAAGGGCTTCTACCCCTCTGACATTGCTGTGGAGTGGGAGA

[0478] GCAATGGGCAGCCAGAGAACAACTACAAGACCACACCTCCTGTCC

[0479] TGGACAGCGACGGCTCCTTCTTCCTCTACAGTAAGCTCACTGTGGA

[0480] CAAGAGCCGCTGGCAGCAGGGCAATGTCTTCTCCTGCAGCGTGAT

[0481] GCACGAGGCCCTGCACAATGCCTATACCCAGAAGTCACTCTCTCT

[0482] GAGCCCTGGAAAGAAGGATCCCAAGTTTTGGGTCCTCGTGGTGGT

[0483] CGGGGGTGTGCTGGCCTGCTACAGCTTGCTGGTCACAGTGGCCTT

[0484] CATCATCTTCTGGGTGCGCTCCAAGAGGAGCCGGCTGCTTCACAG

[0485] TGATTACATGAACATGACCCCCAGGAGGCCAGGACCCACCAGGAA

[0486] GCACTACCAGCCCTACGCTCCCCCGCGGGACTTTGCTGCTTACCG

[0487] CAGCAGGGTCAAATTTTCTAGATCTGCAGATGCGCCGGCCTATCAA caaggccagaaccagctcTATAACGAGCTCAATCTAGGACGAAGAGAGGA

[0488] GTACGATGTTTTGGACAAGAGACGTGGCCGGGACCCTGAGATGGG

[0489] GGGAAAGCCGAGAAGGAAGAACCCTCAGGAAGGCCTGTACAATGA

[0490] ACTGCAGAAAGATAAGATGGCGGAGGCCTACAGTGAGATTGGGAT

[0491] GAAAGGCGAGCGCCGGAGGGGCAAGGGGCACGATGGCCTTTACC

[0492] AGGGTCTCAGTACAGCCACCAAGGACACCTACGACGCCCTTCACA

[0493] T GCAGGCCCT GCCCCCT CGCT AAtaagaattctaactagagctcgctgatcagcct cgactgtgccttctagttgccagccatctgttgtttgcccctcccccgtgccttccttgaccctggaaggt gccactcccactgtcctttcctaataaaatgaggaaattgcatcgcattgtctgagtaggtgtcattcta ttctggggggtggggtggggcaggacagcaagggggaggattgggaagagaatagcaggcat gctggggataagggcagctttggtgccttcgcaggctgtttccttgcttcaggaatggccaggttctgc ccagagctctggtcaatgatgtctaaaactcctctgattggtggtctcggccttatccattgccaccaa aaccctctttttactaagaaacagtgagccttgttctggcagtccagagaatgacacgggaaaaaa gcagatgaagagaaggtggcaggagagggcacgtggcccagcctcagtctctccaactgagttc ctgcctgcctgcctttgctcagactgtttgccccttactgctcttctaggcctcattctaagccccttctcca agttgcctctccttatttctccctgtctgccaaaaaatctttcccagctcactaagtcagtctcacgcagt cactc tgatagcttgtgcctgtccctgagtcccagtccatcacgagcagctggtttctaagatgctatttcccgt ataaagcatgagaccgtgacttgccagccccacagagccccgcccttgtccatcactggcatctgg actccagcctgggttggggcaaagagggaaatgagatcatgtcctaaccctgatcctcttgtcccac agatatccagaaccctgaccctgccgtgtaccagctgagagactctaaatccagtgacaagtctgt ctgcctattcaccgattttgattctcaaacaaatgtgtcacaaagtaaggattctgatgtgtatatcaca gacaaaactgtgctagacatgaggtctGGATCCGGAGCCACCAACTTCAGCCT GCTGAAGCAGGCCGGCGACGTGGAAGAAAATCCTGGGCCCATGG

[0494] 94 CTCTTCCTGTGACTGCCCTTCTGCTGCCCCTGGCTCTCCTGCTCCA HDTR14

[0495] TGCTGCCCGGCCAGACATCCAGATGACACAGACAACCAGCAGCCT

[0496] CAGCGCCAGCCTGGGGGACAGAGTGACCATTAGCTGCCGGGCCT

[0497] CTCAGGACATCAGCAAATACCTGAACTGGTACCAGCAGAAACCAG

[0498] ATGGCACTGTCAAGCTGCTGATTTACCACACCTCCAGGCTCCACAG

[0499] CGGCGTGCCCAGTCGCTTCAGCGGCAGTGGGAGCGGGACAGATT

[0500] ATTCCCTCACAATCTCCAACCTGGAGCAGGAAGATATTGCCACATA

[0501] CTTCTGCCAGCAAGGCAACACCCTGCCATACACATTTGGAGGCGG CACCAAATTGGAGATCACCGGCGGTGGTGGATCTGGAGGAGGGG

[0502] GCAGCGGAGGTGGCGGCTCTGAGGTGAAACTGCAGGAGAGTGGC

[0503] CCTGGCCTGGTGGCTCCCAGCCAGAGCCTTTCTGTCACCTGCACC

[0504] GTGTCTGGGGTGTCCCTGCCTGACTATGGAGTCTCCTGGATCCGG

[0505] CAGCCTCCAAGAAAAGGACTGGAATGGCTGGGCGTCATCTGGGGA

[0506] AGTGAGACCACCTACTATAATTCAGCCCTCAAGTCCCGGCTCACCA

[0507] TCATTAAGGACAACTCCAAATCCCAGGTGTTCCTGAAGATGAATTC

[0508] TCTCCAGACTGATGACACAGCCATCTACTACTGTGCCAAGCATTAT

[0509] TATTACGGCGGGTCCTATGCCATGGACTACTGGGGCCAGGGGACC

[0510] AGTGTCACTGTTTCTTCTGAAAGCAAGTATGGCCCCCCATGCCCTC

[0511] CCTGCCCCGGACAGCCGCGGGAGCCCCAGGTCTACACCCTCCCA

[0512] CCCTCCAGAGATGAGCTGACCAAGAACCAGGTTTCACTGACATGC

[0513] CTGGTGAAGGGCTTCTACCCCTCTGACATTGCTGTGGAGTGGGAG

[0514] AGCAATGGGCAGCCAGAGAACAACTACAAGACCACACCTCCTGTC

[0515] CTGGACAGCGACGGCTCCTTCTTCCTCTACAGTAAGCTCACTGTG

[0516] GACAAGAGCCGCTGGCAGCAGGGCAATGTCTTCTCCTGCAGCGTG

[0517] ATGCACGAGGCCCTGCACAATGCCTATACCCAGAAGTCACTCTCTC

[0518] TGAGCCCTGGAAAGAAGGATCCCAAGTTTTGGGTCCTCGTGGTGG

[0519] TCGGGGGTGTGCTGGCCTGCTACAGCTTGCTGGTCACAGTGGCCT

[0520] TCATCATCTTCTGGGTGCGCTCCAAGAGGAGCCGGCTGCTTCACA

[0521] GTGATTACATGAACATGACCCCCAGGAGGCCAGGACCCACCAGGA

[0522] AGCACTACCAGCCCTACGCTCCCCCGCGGGACTTTGCTGCTTACC

[0523] GCAGCAGGGTCAAATTTTCTAGATCTGCAGATGCGCCGGCCTATC

[0524] AAcaaggccagaaccagctcTATAACGAGCTCAATCTAGGACGAAGAGAG

[0525] GAGTACGATGTTTTGGACAAGAGACGTGGCCGGGACCCTGAGATG

[0526] GGGGGAAAGCCGAGAAGGAAGAACCCTCAGGAAGGCCTGTACAA

[0527] TGAACTGCAGAAAGATAAGATGGCGGAGGCCTACAGTGAGATTGG

[0528] GATGAAAGGCGAGCGCCGGAGGGGCAAGGGGCACGATGGCCTTT

[0529] ACCAGGGTCTCAGTACAGCCACCAAGGACACCTACGACGCCCTTC

[0530] ACAT GCAGGCCCT GCCCCCT CGCT AAtaagaattctaactagagctcgctgatca gcctcgactgtgccttctagttgccagccatctgttgtttgcccctcccccgtgccttccttgaccctgga aggtgccactcccactgtcctttcctaataaaatgaggaaattgcatcgcattgtctgagtaggtgtca ttctattctggggggtggggtggggcaggacagcaagggggaggattgggaagagaatagcag gcatgctggggataagggcagctttggtgccttcgcaggctgtttccttgcttcaggaatggccaggtt ctgcccagagctctggtcaatgatgtctaaaactcctctgattggtggtctcggccttatccattgccac caaaaccctctttttactaagaaacagtgagccttgttctggcagtccagagaatgacacgggaaa aaagcagatgaagagaaggtggcaggagagggcacgtggcccagcctcagtctctccaactga gttcctgcctgcctgcctttgctcagactgtttgccccttactgctcttctaggcctcattctaagccccttc tccaagttgcctctccttatttctccctgtctgccaaaaaatctttcccagctcactaagtcagtctcacg cagtcactc tgatagcttgtgcctgtccctgagtcccagtccatcacgagcagctggtttctaagatgctatttcccgt ataaagcatgagaccgtgacttgccagccccacagagccccgcccttgtccatcactggcatctgg

[0531] 95 HDTR15 actccagcctgggttggggcaaagagggaaatgagatcatgtcctaaccctgatcctcttgtcccac agatatccagaaccctgaccctgccgtgtaccagctgagagactctaaatccagtgacaagtctgt ctgcctattcaccgattttgattctcaaacaaatgtgtcacaaagtaaggattctgatgtgtatatcaca gacaaaactgtgctagacatgaggtctatggacttcaagagcaacagtgctgtgGGATCCGG AGCCACCAACTTCAGCCTGCTGAAGCAGGCCGGCGACGTGGAAG

[0532] AAAATCCTGGGCCCATGGCTCTTCCTGTGACTGCCCTTCTGCTGCC

[0533] CCTGGCTCTCCTGCTCCATGCTGCCCGGCCAGACATCCAGATGAC

[0534] ACAGACAACCAGCAGCCTCAGCGCCAGCCTGGGGGACAGAGTGA

[0535] CCATTAGCTGCCGGGCCTCTCAGGACATCAGCAAATACCTGAACT

[0536] GGTACCAGCAGAAACCAGATGGCACTGTCAAGCTGCTGATTTACC

[0537] ACACCTCCAGGCTCCACAGCGGCGTGCCCAGTCGCTTCAGCGGC

[0538] AGTGGGAGCGGGACAGATTATTCCCTCACAATCTCCAACCTGGAG

[0539] CAGGAAGATATTGCCACATACTTCTGCCAGCAAGGCAACACCCTG

[0540] CCATACACATTTGGAGGCGGCACCAAATTGGAGATCACCGGCGGT

[0541] GGTGGATCTGGAGGAGGGGGCAGCGGAGGTGGCGGCTCTGAGG

[0542] TGAAACTGCAGGAGAGTGGCCCTGGCCTGGTGGCTCCCAGCCAG

[0543] AGCCTTTCTGTCACCTGCACCGTGTCTGGGGTGTCCCTGCCTGAC

[0544] TATGGAGTCTCCTGGATCCGGCAGCCTCCAAGAAAAGGACTGGAA

[0545] TGGCTGGGCGTCATCTGGGGAAGTGAGACCACCTACTATAATTCA

[0546] GCCCTCAAGTCCCGGCTCACCATCATTAAGGACAACTCCAAATCCC

[0547] AGGTGTTCCTGAAGATGAATTCTCTCCAGACTGATGACACAGCCAT

[0548] CTACTACTGTGCCAAGCATTATTATTACGGCGGGTCCTATGCCATG

[0549] GACTACTGGGGCCAGGGGACCAGTGTCACTGTTTCTTCTGAAAGC

[0550] AAGTATGGCCCCCCATGCCCTCCCTGCCCCGGACAGCCGCGGGA

[0551] GCCCCAGGTCTACACCCTCCCACCCTCCAGAGATGAGCTGACCAA

[0552] GAACCAGGTTTCACTGACATGCCTGGTGAAGGGCTTCTACCCCTCT

[0553] GACATTGCTGTGGAGTGGGAGAGCAATGGGCAGCCAGAGAACAAC

[0554] TACAAGACCACACCTCCTGTCCTGGACAGCGACGGCTCCTTCTTC

[0555] CTCTACAGTAAGCTCACTGTGGACAAGAGCCGCTGGCAGCAGGGC

[0556] AATGTCTTCTCCTGCAGCGTGATGCACGAGGCCCTGCACAATGCC

[0557] TATACCCAGAAGTCACTCTCTCTGAGCCCTGGAAAGAAGGATCCCA

[0558] AGTTTTGGGTCCTCGTGGTGGTCGGGGGTGTGCTGGCCTGCTACA

[0559] GCTTGCTGGTCACAGTGGCCTTCATCATCTTCTGGGTGCGCTCCAA

[0560] GAGGAGCCGGCTGCTTCACAGTGATTACATGAACATGACCCCCAG

[0561] GAGGCCAGGACCCACCAGGAAGCACTACCAGCCCTACGCTCCCC

[0562] CGCGGGACTTTGCTGCTTACCGCAGCAGGGTCAAATTTTCTAGATC

[0563] TGCAGATGCGCCGGCCTATCAAcaaggccagaaccagctcTATAACGAGC

[0564] TCAATCTAGGACGAAGAGAGGAGTACGATGTTTTGGACAAGAGAC

[0565] GTGGCCGGGACCCTGAGATGGGGGGAAAGCCGAGAAGGAAGAAC

[0566] CCTCAGGAAGGCCTGTACAATGAACTGCAGAAAGATAAGATGGCG

[0567] GAGGCCTACAGTGAGATTGGGATGAAAGGCGAGCGCCGGAGGGG

[0568] CAAGGGGCACGATGGCCTTTACCAGGGTCTCAGTACAGCCACCAA

[0569] GGACACCTACGACGCCCTTCACATGCAGGCCCTGCCCCCTCGCTA

[0570] Ataagaattctaactagagctcgctgatcagcctcgactgtgccttctagttgccagccatctgttgttt gcccctcccccgtgccttccttgaccctggaaggtgccactcccactgtcctttcctaataaaatgag gaaattgcatcgcattgtctgagtaggtgtcattctattctggggggtggggtggggcaggacagca agggggaggattgggaagagaatagcaggcatgctggggataagggcagctttggtgccttcgc aggctgtttccttgcttcaggaatggccaggttctgcccagagctctggtcaatgatgtctaaaactcc tctgattggtggtctcggccttatccattgccaccaaaaccctctttttactaagaaacagtgagccttgt tctggcagtccagagaatgacacgggaaaaaagcagatgaagagaaggtggcaggagaggg cacgtggcccagcctcagtctctccaactgagttcctgcctgcctgcctttgctcagactgtttgcccct tactgctcttctaggcctcattctaagccccttctccaagttgcctctccttatttctccctgtctgccaaaa aatctttcccagctcactaagtcagtctcacgcagtcactc aagcatgagaccgtgacttgccagccccacagagccccgcccttgtccatcactggcatctggact ccagcctgggttggggcaaagagggaaatgagatcatgtcctaaccctgatcctcttgtcccacag atatccagaaccctgaccctgccgtgtaccagctgagagactctaaatccagtgacaagtctgtctg cctattcaccgattttgattctcaaacaaatgtgtcacaaagtaaggattctgatgtgtatatcacagac aaaactgtgctagacatgaggtctatggacttcaagagcaacagtgctgtggcctggagcaacaa atctgactttgcatgtgcaaacgccttcaacaacagcattattccagaagacaccttcttccccagc GGATCCGGAGCCACCAACTTCAGCCTGCTGAAGCAGGCCGGCGA

[0571] CGTGGAAGAAAATCCTGGGCCCATGGCTCTTCCTGTGACTGCCCT

[0572] TCTGCTGCCCCTGGCTCTCCTGCTCCATGCTGCCCGGCCAGACAT

[0573] CCAGATGACACAGACAACCAGCAGCCTCAGCGCCAGCCTGGGGG

[0574] ACAGAGTGACCATTAGCTGCCGGGCCTCTCAGGACATCAGCAAAT

[0575] ACCTGAACTGGTACCAGCAGAAACCAGATGGCACTGTCAAGCTGC

[0576] TGATTTACCACACCTCCAGGCTCCACAGCGGCGTGCCCAGTCGCT

[0577] TCAGCGGCAGTGGGAGCGGGACAGATTATTCCCTCACAATCTCCA

[0578] ACCTGGAGCAGGAAGATATTGCCACATACTTCTGCCAGCAAGGCA

[0579] ACACCCTGCCATACACATTTGGAGGCGGCACCAAATTGGAGATCA

[0580] CCGGCGGTGGTGGATCTGGAGGAGGGGGCAGCGGAGGTGGCGG

[0581] CTCTGAGGTGAAACTGCAGGAGAGTGGCCCTGGCCTGGTGGCTC

[0582] 96 CCAGCCAGAGCCTTTCTGTCACCTGCACCGTGTCTGGGGTGTCCC HDTR16

[0583] TGCCTGACTATGGAGTCTCCTGGATCCGGCAGCCTCCAAGAAAAG

[0584] GACTGGAATGGCTGGGCGTCATCTGGGGAAGTGAGACCACCTACT

[0585] ATAATTCAGCCCTCAAGTCCCGGCTCACCATCATTAAGGACAACTC

[0586] CAAATCCCAGGTGTTCCTGAAGATGAATTCTCTCCAGACTGATGAC

[0587] ACAGCCATCTACTACTGTGCCAAGCATTATTATTACGGCGGGTCCT

[0588] ATGCCATGGACTACTGGGGCCAGGGGACCAGTGTCACTGTTTCTT

[0589] CTGAAAGCAAGTATGGCCCCCCATGCCCTCCCTGCCCCGGACAGC

[0590] CGCGGGAGCCCCAGGTCTACACCCTCCCACCCTCCAGAGATGAG

[0591] CTGACCAAGAACCAGGTTTCACTGACATGCCTGGTGAAGGGCTTC

[0592] TACCCCTCTGACATTGCTGTGGAGTGGGAGAGCAATGGGCAGCCA

[0593] GAGAACAACTACAAGACCACACCTCCTGTCCTGGACAGCGACGGC

[0594] TCCTTCTTCCTCTACAGTAAGCTCACTGTGGACAAGAGCCGCTGGC

[0595] AGCAGGGCAATGTCTTCTCCTGCAGCGTGATGCACGAGGCCCTGC

[0596] ACAATGCCTATACCCAGAAGTCACTCTCTCTGAGCCCTGGAAAGAA

[0597] GGATCCCAAGTTTTGGGTCCTCGTGGTGGTCGGGGGTGTGCTGGC

[0598] CTGCTACAGCTTGCTGGTCACAGTGGCCTTCATCATCTTCTGGGTG

[0599] CGCTCCAAGAGGAGCCGGCTGCTTCACAGTGATTACATGAACATG

[0600] ACCCCCAGGAGGCCAGGACCCACCAGGAAGCACTACCAGCCCTA CGCTCCCCCGCGGGACTTTGCTGCTTACCGCAGCAGGGTCAAATT

[0601] TT CT AGAT CT GCAGAT GCGCCGGCCT AT CAAcaaggccagaaccagctcT ATAACGAGCTCAATCTAGGACGAAGAGAGGAGTACGATGTTTTGGA

[0602] CAAGAGACGTGGCCGGGACCCTGAGATGGGGGGAAAGCCGAGAA

[0603] GGAAGAACCCTCAGGAAGGCCTGTACAATGAACTGCAGAAAGATA

[0604] AGATGGCGGAGGCCTACAGTGAGATTGGGATGAAAGGCGAGCGC

[0605] CGGAGGGGCAAGGGGCACGATGGCCTTTACCAGGGTCTCAGTAC

[0606] AGCCACCAAGGACACCTACGACGCCCTTCACATGCAGGCCCTGCC

[0607] CCCTCGCTAAtaagaattctaactagagctcgctgatcagcctcgactgtgccttctagttgcc agccatctgttgtttgcccctcccccgtgccttccttgaccctggaaggtgccactcccactgtcctttcc taataaaatgaggaaattgcatcgcattgtctgagtaggtgtcattctattctggggggtggggtggg gcaggacagcaagggggaggattgggaagagaatagcaggcatgctggggataagggcagc tttggtgccttcgcaggctgtttccttgcttcaggaatggccaggttctgcccagagctctggtcaatga tgtctaaaactcctctgattggtggtctcggccttatccattgccaccaaaaccctctttttactaagaaa cagtgagccttgttctggcagtccagagaatgacacgggaaaaaagcagatgaagagaaggtg gcaggagagggcacgtggcccagcctcagtctctccaactgagttcctgcctgcctgcctttgctca gactgtttgccccttactgctcttctaggcctcattctaagccccttctccaagttgcctctccttatttctcc ctgtctgccaaaaaatctttcccagctcactaagtcagtctcacgcagtcactc acaccggggctaacatggctgggcttctgctgcagtggcaggagctctgtgcacagagaacagc ctcatctgctcgccttgtttccacctcccctcccattgccccaggttctttggccccacagcggccacat ctgccgttggtgccaataggttttccaggagctggttgaggtgggagggagggagagggttgtgatc aggctgaggcatggggattggatatagtctccgtgtcatgatttatttggtcagtcagtcctagtgcca ccctggggtaatggggatgtgttctcgtcaccttgggcctggctgaccagctttatctcttggcacaga ggcacagagctttggcctgctggatcTcaaactctgctacctgctggatggaatcctcttcggcagc ggcGCTACCAATTTTTCTTTGTTGAAGCAGGCCGGGGATGTGGAGG AAAATCCGGGGCCAATGGCTCTTCCTGTGACTGCCCTTCTGCTGC

[0608] CCCTGGCTCTCCTGCTCCATGCTGCCCGGCCAGACATCCAGATGA

[0609] CACAGACAACCAGCAGCCTCAGCGCCAGCCTGGGGGACAGAGTG

[0610] ACCATTAGCTGCCGGGCCTCTCAGGACATCAGCAAATACCTGAACT

[0611] GGTACCAGCAGAAACCAGATGGCACTGTCAAGCTGCTGATTTACC

[0612] ACACCTCCAGGCTCCACAGCGGCGTGCCCAGTCGCTTCAGCGGC

[0613] 97 HDTR19

[0614] AGTGGGAGCGGGACAGATTATTCCCTCACAATCTCCAACCTGGAG

[0615] CAGGAAGATATTGCCACATACTTCTGCCAGCAAGGCAACACCCTG

[0616] CCATACACATTTGGAGGCGGCACCAAATTGGAGATCACCGGCGGT

[0617] GGTGGATCTGGAGGAGGGGGCAGCGGAGGTGGCGGCTCTGAGG

[0618] TGAAACTGCAGGAGAGTGGCCCTGGCCTGGTGGCTCCCAGCCAG

[0619] AGCCTTTCTGTCACCTGCACCGTGTCTGGGGTGTCCCTGCCTGAC

[0620] TATGGAGTCTCCTGGATCCGGCAGCCTCCAAGAAAAGGACTGGAA

[0621] TGGCTGGGCGTCATCTGGGGAAGTGAGACCACCTACTATAATTCA

[0622] GCCCTCAAGTCCCGGCTCACCATCATTAAGGACAACTCCAAATCCC

[0623] AGGTGTTCCTGAAGATGAATTCTCTCCAGACTGATGACACAGCCAT

[0624] CTACTACTGTGCCAAGCATTATTATTACGGCGGGTCCTATGCCATG

[0625] GACTACTGGGGCCAGGGGACCAGTGTCACTGTTTCTTCTGAAAGC

[0626] AAGTATGGCCCCCCATGCCCTCCCTGCCCCGGACAGCCGCGGGA GCCCCAGGTCTACACCCTCCCACCCTCCAGAGATGAGCTGACCAA

[0627] GAACCAGGTTTCACTGACATGCCTGGTGAAGGGCTTCTACCCCTCT

[0628] GACATTGCTGTGGAGTGGGAGAGCAATGGGCAGCCAGAGAACAAC

[0629] TACAAGACCACACCTCCTGTCCTGGACAGCGACGGCTCCTTCTTC

[0630] CTCTACAGTAAGCTCACTGTGGACAAGAGCCGCTGGCAGCAGGGC

[0631] AATGTCTTCTCCTGCAGCGTGATGCACGAGGCCCTGCACAATGCC

[0632] TATACCCAGAAGTCACTCTCTCTGAGCCCTGGAAAGAAGGATCCCA

[0633] AGTTTTGGGTCCTCGTGGTGGTCGGGGGTGTGCTGGCCTGCTACA

[0634] GCTTGCTGGTCACAGTGGCCTTCATCATCTTCTGGGTGCGCTCCAA

[0635] GAGGAGCCGGCTGCTTCACAGTGATTACATGAACATGACCCCCAG

[0636] GAGGCCAGGACCCACCAGGAAGCACTACCAGCCCTACGCTCCCC

[0637] CGCGGGACTTTGCTGCTTACCGCAGCagagtgaaggtgggtaccactgggcttt gggaggagggcacggggtcccccacttgatggatgttcagaggggccttggtcttggaaggtctca agctcgggtggtgcctggggcttggtatccaggagcaaagcaaggaccagccaagtgtgtgcctt gagtgggctgaggaggaggtggcagtgtctggctgagatggacagggtaggagggagagcctg gtgctaggcacctccatgacaagccgtacaaatgtgtgcacatcagagtgtcccagggaaggcg atgctactggtgacaaaggggcttacactcaggcagaggtccttctttccaagtgtgaatgaaggcc atgttagcctttctcttgaaaaggccctttcctcatctgtaactggggaG acaccggggctaacatggctgggcttctgctgcagtggcaggagctctgtgcacagagaacagc ctcatctgctcgccttgtttccacctcccctcccattgccccaggttctttggccccacagcggccacat ctgccgttggtgccaataggttttccaggagctggttgaggtgggagggagggagagggttgtgatc aggctgaggcatggggattggatatagtctccgtgtcatgatttatttggtcagtcagtcctagtgcca ccctggggtaatggggatgtgttctcgtcaccttgggcctggctgaccagctttatctcttggcacaga ggcacagagctttggcctgctggatcTcaaactctgctacctgctggatggaatcctcttcggcagc ggcGCTACCAATTTTTCTTTGTTGAAGCAGGCCGGGGATGTGGAGG AAAATCCGGGGCCAATGGCTCTTCCTGTGACTGCCCTTCTGCTGC

[0638] CCCTGGCTCTCCTGCTCCATGCTGCCCGGCCAGACATCCAGATGA

[0639] CACAGACAACCAGCAGCCTCAGCGCCAGCCTGGGGGACAGAGTG

[0640] ACCATTAGCTGCCGGGCCTCTCAGGACATCAGCAAATACCTGAACT

[0641] GGTACCAGCAGAAACCAGATGGCACTGTCAAGCTGCTGATTTACC

[0642] ACACCTCCAGGCTCCACAGCGGCGTGCCCAGTCGCTTCAGCGGC

[0643] 98 HDTR25

[0644] AGTGGGAGCGGGACAGATTATTCCCTCACAATCTCCAACCTGGAG

[0645] CAGGAAGATATTGCCACATACTTCTGCCAGCAAGGCAACACCCTG

[0646] CCATACACATTTGGAGGCGGCACCAAATTGGAGATCACCGGCGGT

[0647] GGTGGATCTGGAGGAGGGGGCAGCGGAGGTGGCGGCTCTGAGG

[0648] TGAAACTGCAGGAGAGTGGCCCTGGCCTGGTGGCTCCCAGCCAG

[0649] AGCCTTTCTGTCACCTGCACCGTGTCTGGGGTGTCCCTGCCTGAC

[0650] TATGGAGTCTCCTGGATCCGGCAGCCTCCAAGAAAAGGACTGGAA

[0651] TGGCTGGGCGTCATCTGGGGAAGTGAGACCACCTACTATAATTCA

[0652] GCCCTCAAGTCCCGGCTCACCATCATTAAGGACAACTCCAAATCCC

[0653] AGGTGTTCCTGAAGATGAATTCTCTCCAGACTGATGACACAGCCAT

[0654] CTACTACTGTGCCAAGCATTATTATTACGGCGGGTCCTATGCCATG

[0655] GACTACTGGGGCCAGGGGACCAGTGTCACTGTTTCTTCTGAAAGC

[0656] AAGTATGGCCCCCCATGCCCTCCCTGCCCCGGACAGCCGCGGGA GCCCCAGGTCTACACCCTCCCACCCTCCAGAGATGAGCTGACCAA

[0657] GAACCAGGTTTCACTGACATGCCTGGTGAAGGGCTTCTACCCCTCT

[0658] GACATTGCTGTGGAGTGGGAGAGCAATGGGCAGCCAGAGAACAAC

[0659] TACAAGACCACACCTCCTGTCCTGGACAGCGACGGCTCCTTCTTC

[0660] CTCTACAGTAAGCTCACTGTGGACAAGAGCCGCTGGCAGCAGGGC

[0661] AATGTCTTCTCCTGCAGCGTGATGCACGAGGCCCTGCACAATGCC

[0662] TATACCCAGAAGTCACTCTCTCTGAGCCCTGGAAAGAAGGATCCCA

[0663] AGTTTTGGGTCCTCGTGGTGGTCGGGGGTGTGCTGGCCTGCTACA

[0664] GCTTGCTGGTCACAGTGGCCTTCATCATCTTCTGGGTGCGCTCCAA

[0665] GAGGAGCCGGCTGCTTCACAGTGATTACATGAACATGACCCCCAG

[0666] GAGGCCAGGACCCACCAGGAAGCACTACCAGCCCTACGCTCCCC

[0667] CGCGGGACTTTGCTGCTTACCGCAGCagagtgaaggtgggtaccactgCgcttt gCgaggagggcacggggtcccccacttgatggatgttcagaggggccttggtcttggaaggtctc aagctcgggtggtgcctggggcttggtatccaggagcaaagcaaggaccagccaagtgtgtgcct tgagtgggctgaggaggaggtggcagtgtctggctgagatggacagggtaggagggagagcct ggtgctaggcacctccatgacaagccgtacaaatgtgtgcacatcagagtgtcccagggaaggc gatgctactggtgacaaaggggcttacactcaggcagaggtccttctttccaagtgtgaatgaaggc catgttagcctttctcttgaaaaggccctttcctcatctgtaactggggaG acaccggggctaacatggctgggcttctgctgcagtggcaggagctctgtgcacagagaacagc ctcatctgctcgccttgtttccacctcccctcccattgccccaggttctttggccccacagcggccacat ctgccgttggtgccaataggttttccaggagctggttgaggtgggagggagggagagggttgtgatc aggctgaggcatggggattggatatagtctccgtgtcatgatttatttggtcagtcagtcctagtgcca ccctggggtaatggggatgtgttctcgtcaccttgggcctggctgaccagctttatctcttggcacaga ggcacagagctttggcctgctggatcccaaactctgctacctgctggatggaatcctcttcggcagc ggcGCTACCAATTTTTCTTTGTTGAAGCAGGCCGGGGATGTGGAGG AAAATCCGGGGCCAATGGCTCTTCCTGTGACTGCCCTTCTGCTGC

[0668] CCCTGGCTCTCCTGCTCCATGCTGCCCGGCCAGACATCCAGATGA

[0669] CACAGACAACCAGCAGCCTCAGCGCCAGCCTGGGGGACAGAGTG

[0670] ACCATTAGCTGCCGGGCCTCTCAGGACATCAGCAAATACCTGAACT

[0671] GGTACCAGCAGAAACCAGATGGCACTGTCAAGCTGCTGATTTACC

[0672] ACACCTCCAGGCTCCACAGCGGCGTGCCCAGTCGCTTCAGCGGC

[0673] 99 HDTR28

[0674] AGTGGGAGCGGGACAGATTATTCCCTCACAATCTCCAACCTGGAG

[0675] CAGGAAGATATTGCCACATACTTCTGCCAGCAAGGCAACACCCTG

[0676] CCATACACATTTGGAGGCGGCACCAAATTGGAGATCACCGGCGGT

[0677] GGTGGATCTGGAGGAGGGGGCAGCGGAGGTGGCGGCTCTGAGG

[0678] TGAAACTGCAGGAGAGTGGCCCTGGCCTGGTGGCTCCCAGCCAG

[0679] AGCCTTTCTGTCACCTGCACCGTGTCTGGGGTGTCCCTGCCTGAC

[0680] TATGGAGTCTCCTGGATCCGGCAGCCTCCAAGAAAAGGACTGGAA

[0681] TGGCTGGGCGTCATCTGGGGAAGTGAGACCACCTACTATAATTCA

[0682] GCCCTCAAGTCCCGGCTCACCATCATTAAGGACAACTCCAAATCCC

[0683] AGGTGTTCCTGAAGATGAATTCTCTCCAGACTGATGACACAGCCAT

[0684] CTACTACTGTGCCAAGCATTATTATTACGGCGGGTCCTATGCCATG

[0685] GACTACTGGGGCCAGGGGACCAGTGTCACTGTTTCTTCTGAAAGC

[0686] AAGTATGGCCCCCCATGCCCTCCCTGCCCCGGACAGCCGCGGGA GCCCCAGGTCTACACCCTCCCACCCTCCAGAGATGAGCTGACCAA

[0687] GAACCAGGTTTCACTGACATGCCTGGTGAAGGGCTTCTACCCCTCT

[0688] GACATTGCTGTGGAGTGGGAGAGCAATGGGCAGCCAGAGAACAAC

[0689] TACAAGACCACACCTCCTGTCCTGGACAGCGACGGCTCCTTCTTC

[0690] CTCTACAGTAAGCTCACTGTGGACAAGAGCCGCTGGCAGCAGGGC

[0691] AATGTCTTCTCCTGCAGCGTGATGCACGAGGCCCTGCACAATGCC

[0692] TATACCCAGAAGTCACTCTCTCTGAGCCCTGGAAAGAAGGATCCCA

[0693] AGTTTTGGGTCCTCGTGGTGGTCGGGGGTGTGCTGGCCTGCTACA

[0694] GCTTGCTGGTCACAGTGGCCTTCATCATCTTCTGGGTGCGCTCCAA

[0695] GAGGAGCCGGCTGCTTCACAGTGATTACATGAACATGACCCCCAG

[0696] GAGGCCAGGACCCACCAGGAAGCACTACCAGCCCTACGCTCCCC

[0697] CGCGGGACTTTGCTGCTTACCGCAGCagagtgaaggtgggtaccactgCgcttt gCgaggagggcacggggtcccccacttgatggatgttcagaggggccttggtcttggaaggtctc aagctcgggtggtgcctggggcttggtatccaggagcaaagcaaggaccagccaagtgtgtgcct tgagtgggctgaggaggaggtggcagtgtctggctgagatggacagggtaggagggagagcct ggtgctaggcacctccatgacaagccgtacaaatgtgtgcacatcagagtgtcccagggaaggc gatgctactggtgacaaaggggcttacactcaggcagaggtccttctttccaagtgtgaatgaaggc catgttagcctttctcttgaaaaggccctttcctcatctgtaactggggaG gtactccaaagattcaggtttactcacgtcatccagcagagaatggaaagtcaaatttcctgaattgc tatgtgtctgggtttcatccatccgacattgaagttgacttactgaagaatggagagagaattgaaaa agtggagcattcagacttgtctttcagcaaggactggtctttctatctcttgtactacactgaattcaccc ccactgaaaaagatgagtatgcAT GCCGT GT GAACCAT GT GACTTT GT CACA GCCTAAAATTGTGAAATGGGACAGGGACATGGGTGGGGGCGGATC

[0698] AGGCGGCGGCGGTTCTGGGGGAGGCGGCTCAGGTGGCGGTGGG

[0699] AGTGGATCCCATTCCCTGAAGTATTTCCACACCAGCGTTAGTCGGC

[0700] CGGGAAGGGGAGAACCAAGATTCATTTCCGTCGGCTATGTCGACG

[0701] ATACCCAATTTGTGCGATTTGATAATGACGCAGCTTCACCCCGCAT

[0702] GGTGCCTCGGGCTCCTTGGATGGAGCAAGAAGGCTCAGAGTACTG

[0703] GGACCGGGAGACCCGATCTGCGCGCGATACAGCACAAATCTTTAG

[0704] GGTCAACCTTCGAACATTGAGGGGCTACTACAACCAGAGTGAGGC

[0705] AGGTTCCCATACGTTGCAATGGATGCATGGTTGCGAACTTGGTCCA

[0706] 100 HDTR30

[0707] GATGGGAGGTTCCTCAGAGGTTATGAACAATTTGCTTACGATGGAA

[0708] AGGACTACCTTACACTCAATGAGGACCTCCGCAGCTGGACCGCCG

[0709] TTGACACCGCTGCTCAAATCTCCGAGCAAAAGAGTAATGACGCGTC

[0710] AGAAGCGGAGCATCAGCGCGCCTACCTCGAGGACACGTGCGTTG

[0711] AGTGGCTCCATAAATACCTGGAAAAAGGGAAAGAAACCCTCCTCCA

[0712] CCTGGAGCCACCGAAGACCCACGTCACGCATCACCCAATATCCGA

[0713] TCACGAGGCTACACTGAGGTGCTGGGCTCTCGGTTTCTATCCGGC

[0714] AGAAATAACCCTGACCTGGCAGCAGGATGGTGAGGGCCATACGCA

[0715] GGACACGGAGCTTGTCGAGACTCGACCCGCCGGTGATGGAACATT

[0716] CCAAAAATGGGCTGCTGTCGTTGTACCCTCTGGAGAAGAACAACG

[0717] ATACACCTGCCACGTGCAGCATGAGGGACTGCCGGAACCGGTAAC

[0718] CTTGCGCTGGAAGCCGGCGTCACAACCAACGATACCCATCGTCGG

[0719] CATTATAGCCGGACTTGTTCTCCTGGGAAGCGTGGTCAGCGGAGC TGTCGTAGCGGCCGTAATTTGGAGGAAAAAGTCTTCTGGGGGAAA

[0720] GGGTGGGTCATACTCTAAAGCGGAGTGGTCTGACTCAGCACAGGG

[0721] TTCCGAATCCCACTCCCTGTGAGCGGCCGCGTCGAGTCTAGAGGG

[0722] CCCGTTTAAACCCGCTGATCAGCCTCGACTGTGCCTTCTAGTTGCC

[0723] AGCCATCTGTTGTTTGCCCCTCCCCCGTGCCTTCCTTGACCCTGGA

[0724] AGGTGCCACTCCCACTGTCCTTTCCTAATAAAATGAGGAAATTGCA

[0725] TCGCATTGTCTGAGTAGGTGTCATTCTATTCTGGGGGGTGGGGTG

[0726] GGGCAGGACAGCAAGGGGGAGGATTGGGAAGACAATAGCAGGCA

[0727] TGCT GGGGAT GCGGT GGGCT CT AT GGgtcagggaatgttcttaaagatcagatt agtggcacctgctgagatactgatgcacagcatggtttctgaaccagtagtttccctgcagttgagca gggagcagcagcagcacttgcacaaatacatatacactcttaacacttcttacctactggcttcctcT AGCTTTTGTGGCAGCTTCAGGTATATTTAGCACTGAACGAACATCT

[0728] CAAGAAGGTATAGGCCTTTGTTTGTAAGTCCTGCTGTCCTAGCATC

[0729] CTATAATCCTGGACTTCTCCAGTACTTTCTGGCTGGATTGGTATCT

[0730] GAGGCTAGTAGGA aagctcatttggccagagtggaaatggaattgggagaaatcgatgaccaaatgtaaacacttggtg cctgatatagcttgacaccaagttagccccaagtgaaataccctggcaatattaatgtgtcttttcccg atattcctcaggtactccaaagattcaggtttactcacgtcatccagcagagaatggaaagtcaaatt tcctgaattgctatgtgtctgggtttcatccatccgacattgaagttgacttactgaagaatggagaga gaattgaaaaagtggagcattcagacttgtctttcagcaaggactggtctttctatctcttgtactacact gaattcacccccactgaaaaagatgagtatgctTGCAGGGTTAACCATGTCACATT

[0731] GTCACAACCTAAAATTGTGAAATGGGACAGGGACATGGGTGGGGG

[0732] CGGATCAGGCGGCGGCGGTTCTGGGGGAGGCGGCTCAGGTGGC

[0733] GGTGGGAGTGGATCCCATTCCCTGAAGTATTTCCACACCAGCGTTA

[0734] GTCGGCCGGGAAGGGGAGAACCAAGATTCATTTCCGTCGGCTATG

[0735] TCGACGATACCCAATTTGTGCGATTTGATAATGACGCAGCTTCACC

[0736] CCGCATGGTGCCTCGGGCTCCTTGGATGGAGCAAGAAGGCTCAGA

[0737] GTACTGGGACCGGGAGACCCGATCTGCGCGCGATACAGCACAAAT

[0738] CTTTAGGGTCAACCTTCGAACATTGAGGGGCTACTACAACCAGAGT

[0739] 101 GAGGCAGGTTCCCATACGTTGCAATGGATGCATGGTTGCGAACTT HDTR31

[0740] GGTCCAGATGGGAGGTTCCTCAGAGGTTATGAACAATTTGCTTACG

[0741] ATGGAAAGGACTACCTTACACTCAATGAGGACCTCCGCAGCTGGA

[0742] CCGCCGTTGACACCGCTGCTCAAATCTCCGAGCAAAAGAGTAATG

[0743] ACGCGTCAGAAGCGGAGCATCAGCGCGCCTACCTCGAGGACACG

[0744] TGCGTTGAGTGGCTCCATAAATACCTGGAAAAAGGGAAAGAAACC

[0745] CTCCTCCACCTGGAGCCACCGAAGACCCACGTCACGCATCACCCA

[0746] ATATCCGATCACGAGGCTACACTGAGGTGCTGGGCTCTCGGTTTC

[0747] TATCCGGCAGAAATAACCCTGACCTGGCAGCAGGATGGTGAGGGC

[0748] CATACGCAGGACACGGAGCTTGTCGAGACTCGACCCGCCGGTGAT

[0749] GGAACATTCCAAAAATGGGCTGCTGTCGTTGTACCCTCTGGAGAA

[0750] GAACAACGATACACCTGCCACGTGCAGCATGAGGGACTGCCGGAA

[0751] CCGGTAACCTTGCGCTGGAAGCCGGCGTCACAACCAACGATACCC

[0752] ATCGTCGGCATTATAGCCGGACTTGTTCTCCTGGGAAGCGTGGTC

[0753] AGCGGAGCTGTCGTAGCGGCCGTAATTTGGAGGAAAAAGTCTTCT GGGGGAAAGGGTGGGTCATACTCTAAAGCGGAGTGGTCTGACTCA

[0754] GCACAGGGTTCCGAATCCCACTCCCTGTGACTGTGCCTTCTAGTTG

[0755] CCAGCCATCTGTTGTTTGCCCCTCCCCCGTGCCTTCCTTGACCCTG

[0756] GAAGGTGCCACTCCCACTGTCCTTTCCTAATAAAATGAGGAAATTG

[0757] CATCgtgtgaaccatgtgactttgtcacagcccaagatagttaagtggggtaagtcttacattctttt gtaagctgctgaaagttgtgtatgagtagtcatatcataaagctgctttgatataaaaaaggtctatgg ccatactaccctgaatgagtcccatcccatctgatataaacaatctgcatattgggattgtcagggaa tgttcttaaagatcagattagtggcacctgctgagatactgatgcacagcatggtttctgaaccagta gtttccctgcagttgagcagggagcagcagcagcacttgcacaaatacatatacactcttaacactt cttacctactggcttcctctagcttttgtggcagcttcaggta

[0758] TACCCAACCAAGAGCGTGTTTTTCTTCCACAGACACCAATGTTCAA

[0759] AATGGAGGCTTGGGGGCAAAATTCTTTTGCTATGTCTCTAGTCGTC

[0760] CAAAAAATGGTCCTAACTTTTTCTGACTCCTGCTTGTCAAAAATTGT

[0761] GGGCTCATAGTTAATGCTAGATGCTTCCTTCCTCTATTTCCCCCCA

[0762] AATTTCCTGGGAACCCCTGGTCAATACCAGCAGTAAGTTCCACTGT

[0763] TCTAGGGTGTAGAAATGGCTGTGACGCAGCAGCAAGAGGGAAGGA

[0764] CATCAGATGTCATCAGTGGTCATACTGCAACACAGCGCTTTTTCTG

[0765] TTTAGGAATGCAGGTACCCACAACATTTACTAACACTTTTTTTTTCTT

[0766] ATTTATTTTCTAGTTGGCGTTTGGGGGCAAGAGCAGAAGCTCATCA

[0767] GTGAAGAAGATCTGCAAGTGCAACTGGTGCAGTCAGGAGGAGGAG

[0768] TTGTGCAGCCAGGAGGAAGTCTGCGGGTCAGTTGCGCCGCGTCA

[0769] GGCGTTACTCTCTCCGATTATGGAATGCATTGGGTCAGACAGGCA

[0770] CCAGGCAAAGGTCTGGAATGGATGGCATTCATTCGCAATGACGGG

[0771] TCGGATAAGTATTACGCTGACAGCGTAAAGGGCAGGTTTACAATCT

[0772] CTCGGGATAATAGCAAAAAAACTGTCAGCCTGCAGATGAGTTCACT

[0773] GCGCGCCGAGGACACCGCTGTGTACTATTGTGCTAAAAACGGCGA

[0774] GTCTGGGCCATTGGATTATTGGTACTTTGATCTCTGGGGGCGGGG

[0775] 102 HDTR32

[0776] AACACTGGTGACTGTGTCTGGAGGTGGTGGATCAGGAGGTGGAG

[0777] GTTCCGGAGGTGGAGGAAGTGACGTCGTGATGACTCAGTCTCCCT

[0778] CCTCTCTGAGTGCCAGCGTTGGGGATCGCGTTACGATCACCTGCC

[0779] AAGCCAGCCAGGATATCAGCAACTACTTAAATTGGTATCAGCAAAA

[0780] GCCTGGAAAGGCTCCAAAGCTGCTAATCTACGACGCGTCCAACCT

[0781] GGAGACCGGTGTTCCATCACGCTTCTCTGGATCTGGGTCTGGTAC

[0782] CGATTTTACATTCACCATATCTTCTTTGCAACCCGAGGACTTCGCTA

[0783] CCTATTATTGCCAACAGTATAGCAGCTTCCCTTTAACTTTCGGGGG

[0784] AGGCACTAAAGTGGATATCAAGCGGGGAGGCGGAGGTAGCGGAG

[0785] GCGGCGGAAGCGGTGGAGGTGGGTCTCAAGATGGTGAGATATGC

[0786] TTTCTTTCTTTCTTTTTTATGAAATCACCCCATCATTCTTTGTAGTTA

[0787] TGAATCGAGCTTTCTCTTAGGCCTCCCACAGAACTTCCACAGAGGT

[0788] CAGGAAAAGGAGTTTCTGCCATCTACCCCTTTGACTTTCCTCACAA

[0789] GTCTGGAGATATTTCTAGCCCAGAAGAGGGAAGCAACAGAGGCAG

[0790] GAAATAATGAGTCTTAACCATACAAAAGAAAAATTGAGACTTAAATG

[0791] AAGTTGAAAGCACTAACAGTTTTCATTTGTTTGCATTTCATATTTGAT

[0792] GTGAGATTCTGCAGAGGAGACGTAGCCAGAATGCATGCACAGGGT TACTCTGGATAAGCTGCTGGGGCAACATTTGGATGTGTGTTCAGAA

[0793] TCACATGTCTGAAT

[0794] TACCCAACCAAGAGCGTGTTTTTCTTCCACAGACACCAATGTTCAA

[0795] AATGGAGGCTTGGGGGCAAAATTCTTTTGCTATGTCTCTAGTCGTC

[0796] CAAAAAATGGTCCTAACTTTTTCTGACTCCTGCTTGTCAAAAATTGT

[0797] GGGCTCATAGTTAATGCTAGATGCTTCCTTCCTCTATTTCCCCCCA

[0798] AATTTCCTGGGAACCCCTGGTCAATACCAGCAGTAAGTTCCACTGT

[0799] TCTAGGGTGTAGAAATGGCTGTGACCCAGCAGCAAGAGGGAAGGA

[0800] CATCAGATGTCATCAGTGGTCATACTGCAACACAGCCCTTTTTCTG

[0801] TTTAGGAATGCAGGTACGCACAACATTTACTAACACTTTTTTTTTCT

[0802] TATTTATTTTCTAGTTGGCGTTTGGGGGCAAGAGCAGAAGCTCATC

[0803] AGTGAAGAAGATCTGCAAGTGCAACTGGTGCAGTCAGGAGGAGGA

[0804] GTTGTGCAGCCAGGAGGAAGTCTGCGGGTCAGTTGCGCCGCGTC

[0805] AGGCGTTACTCTCTCCGATTATGGAATGCATTGGGTCAGACAGGCA

[0806] CCAGGCAAAGGTCTGGAATGGATGGCATTCATTCGCAATGACGGG

[0807] TCGGATAAGTATTACGCTGACAGCGTAAAGGGCAGGTTTACAATCT

[0808] CTCGGGATAATAGCAAAAAAACTGTCAGCCTGCAGATGAGTTCACT

[0809] GCGCGCCGAGGACACCGCTGTGTACTATTGTGCTAAAAACGGCGA

[0810] GTCTGGGCCATTGGATTATTGGTACTTTGATCTCTGGGGGCGGGG

[0811] AACACTGGTGACTGTGTCTGGAGGTGGTGGATCAGGAGGTGGAG

[0812] 103 HDTR33

[0813] GTTCCGGAGGTGGAGGAAGTGACGTCGTGATGACTCAGTCTCCCT

[0814] CCTCTCTGAGTGCCAGCGTTGGGGATCGCGTTACGATCACCTGCC

[0815] AAGCCAGCCAGGATATCAGCAACTACTTAAATTGGTATCAGCAAAA

[0816] GCCTGGAAAGGCTCCAAAGCTGCTAATCTACGACGCGTCCAACCT

[0817] GGAGACCGGTGTTCCATCACGCTTCTCTGGATCTGGGTCTGGTAC

[0818] CGATTTTACATTCACCATATCTTCTTTGCAACCCGAGGACTTCGCTA

[0819] CCTATTATTGCCAACAGTATAGCAGCTTCCCTTTAACTTTCGGGGG

[0820] AGGCACTAAAGTGGATATCAAGCGGGGAGGCGGAGGTAGCGGAG

[0821] GCGGCGGAAGCGGTGGAGGTGGGTCTCAAGATGGTGAGATATGC

[0822] TTTCTTTCTTTCTTTTTTATGAAATCACCCCATCATTCTTTGTAGTTA

[0823] TGAATGGAGCTTTCTCTTAGGCCTCCCACAGAACTTCCACAGAGGT

[0824] CAGGAAAAGGAGTTTCTGCCATCTACCCCTTTGACTTTCCTCACAA

[0825] GTCTGGAGATATTTCTAGCCCAGAAGAGGGAAGCAACAGAGGCAG

[0826] GAAATAATGAGTCTTAACCATACAAAAGAAAAATTGAGACTTAAATG

[0827] AAGTTGAAAGCACTAACAGTTTTCATTTGTTTGCATTTCATATTTGAT

[0828] GTGAGATTCTGCAGAGGAGACGTAGCCAGAATGCATGCACAGGGT

[0829] TACTCTGGATAAGCTGCTGGGGCAACATTTGGATGTGTGTTCAGAA

[0830] TCACATGTCTGAAT

[0831] Sequenc

[0832] 104 GGACGACTAGTGGCTCAGAGTATATCAGCTGGTACACGG e Fig.1 A 1 Sequenc

[0833] 105 CCTGCTGATCACCGAGTCTCAGAGAGTCGACCATGTGCC e Fig.1 A 2

[0834] Sequenc

[0835] 106 GGACGACTAGTGGCTCTCTCTCTCTCAGCTGGTACACGG e Fig.1 A 3

[0836] Sequenc

[0837] 107 CCTGCTGATCACCGAGAGAGAGAGAGTCGACCATGTGCC e Fig.1 A 4

[0838] B2

[0839] 108 CTTGCCCCACTTAACTATCT retargetin g sgRNA

[0840] In one preferred embodiment, guide RNA T1 (SEQ ID NO. 1) is used with guide RNA T7 (SEQ ID NO. 7) and HDRT1 (SEQ ID NO. 38). In preferred embodiments, guide RNA T1 (SEQ ID NO. 1) is used with guide RNA T7 (SEQ ID NO. 7) and the homology arms of HDRT1 (SEQ ID NO. 38), without the knocked-in sequence of SEQ ID NO. 38.

[0841] In one preferred embodiment, guide RNA T1 (SEQ ID NO. 1) is used with guide RNA T8 (SEQ ID NO. 8) and HDRT2 (SEQ ID NO. 39). In preferred embodiments, guide RNA T1 (SEQ ID NO. 1) is used with guide RNA T8 (SEQ ID NO. 8) and the homology arms of HDRT2 (SEQ ID NO. 39), without the knocked-in sequence of SEQ ID NO. 39.

[0842] In one preferred embodiment, guide RNA T1 (SEQ ID NO. 1) is used with guide RNA T9 (SEQ ID NO. 9) and HDRT3 (SEQ ID NO. 40). In preferred embodiments, guide RNA T1 (SEQ ID NO. 1) is used with guide RNA T8 (SEQ ID NO. 9) and the homology arms of HDRT3 (SEQ ID NO. 40) without the knocked-in sequence of SEQ ID NO. 40.

[0843] In one preferred embodiment, guide RNA T1 (SEQ ID NO. 1) is used with guide RNA T10 (SEQ ID NO. 10) and HDRT4 (SEQ ID NO. 41). In preferred embodiments, guide RNA T1 (SEQ ID NO. 1) is used with guide RNA T10 (SEQ ID NO. 10) and the homology arms of HDRT4 (SEQ ID NO. 41) without the knocked-in sequence of SEQ ID NO. 41.

[0844] In one preferred embodiment, guide RNA T1 (SEQ ID NO. 1) is used with guide RNA T11 (SEQ ID NO. 11) and HDRT1 (SEQ ID NO. 42). In preferred embodiments, guide RNA T1 (SEQ ID NO. 1) is used with guide RNA T11 (SEQ ID NO. 11) and the homology arms of HDRT1 (SEQ ID NO. 42) without the knocked-in sequence of SEQ ID NO. 42.

[0845] In one preferred embodiment, guide RNA T1 (SEQ ID NO. 1) is used with guide RNA T12 (SEQ ID NO. 12) and HDRT1 (SEQ ID NO. 43). In preferred embodiments, guide RNA T1 (SEQ ID NO. 1) is used with guide RNA T12 (SEQ ID NO. 12) and the homology arms of HDRT1 (SEQ ID NO. 43) without the knocked-in sequence of SEQ ID NO. 43. In one preferred embodiment, guide RNA T2 (SEQ ID NO. 2) is used with guide RNA T8 (SEQ ID NO. 8) and HDRT6 (SEQ ID NO. 44). In preferred embodiments, guide RNA T2 (SEQ ID NO. 2) is used with guide RNA T8 (SEQ ID NO. 8) and the homology arms of HDRT6 (SEQ ID NO. 44) without the knocked-in sequence of SEQ ID NO. 44.

[0846] In one preferred embodiment, guide RNA T2 (SEQ ID NO. 2) is used with guide RNA T9 (SEQ ID NO. 9) and HDRT7 (SEQ ID NO. 45). In preferred embodiments, guide RNA T2 (SEQ ID NO. 2) is used with guide RNA T9 (SEQ ID NO. 9) and the homology arms of HDRT7 (SEQ ID NO. 45) without the knocked-in sequence of SEQ ID NO. 45.

[0847] In one preferred embodiment, guide RNA T2 (SEQ ID NO. 2) is used with guide RNA T 10 (SEQ ID NO. 10) and HDRT8 (SEQ ID NO. 46). In preferred embodiments, guide RNA T2 (SEQ ID NO. 2) is used with guide RNA T10 (SEQ ID NO. 10) and the homology arms of HDRT8 (SEQ ID NO. 46) without the knocked-in sequence of SEQ ID NO. 46.

[0848] In one preferred embodiment, guide RNA T3 (SEQ ID NO. 3) is used with guide RNA T12 (SEQ ID NO. 12) and HDRT9 (SEQ ID NO. 47). In preferred embodiments, guide RNA T3 (SEQ ID NO. 3) is used with guide RNA T12 (SEQ ID NO. 12) and the homology arms of HDRT9 (SEQ ID NO. 47) without the knocked-in sequence of SEQ ID NO. 47.

[0849] In one preferred embodiment, guide RNA Z2 (SEQ ID NO. 14) is used with guide RNA Z5 (SEQ ID NO. 17) and HDRT17 (SEQ ID NO. 48). In preferred embodiments, guide RNA Z2 (SEQ ID NO. 14) is used with guide RNA Z5 (SEQ ID NO. 17) and the homology arms of HDRT17 (SEQ ID NO. 47) without the knocked-in sequence of SEQ ID NO. 48.

[0850] In one preferred embodiment, guide RNA Z2 (SEQ ID NO. 14) is used with guide RNA Z6 (SEQ ID NO. 18) and HDRT17 (SEQ ID NO. 49). In preferred embodiments, guide RNA Z2 (SEQ ID NO. 14) is used with guide RNA Z6 (SEQ ID NO. 18) and the homology arms of HDRT17 (SEQ ID NO. 49) without the knocked-in sequence of SEQ ID NO. 49.

[0851] In one preferred embodiment, guide RNA Z3 (SEQ ID NO. 15) is used with guide RNA Z5 (SEQ ID NO. 17) and HDRT20 (SEQ ID NO. 50). In preferred embodiments, guide RNA Z3 (SEQ ID NO. 15) is used with guide RNA Z5 (SEQ ID NO. 17) and the homology arms of HDRT20 (SEQ ID NO. 50) without the knocked-in sequence of SEQ ID NO. 50.

[0852] In one preferred embodiment, guide RNA Z3 (SEQ ID NO. 15) is used with guide RNA Z6 (SEQ ID NO. 18) and HDRT20 (SEQ ID NO. 51). In preferred embodiments, guide RNA Z3 (SEQ ID NO. 15) is used with guide RNA Z6 (SEQ ID NO. 18) and the homology arms of HDRT20 (SEQ ID NO. 51) without the knocked-in sequence of SEQ ID NO. 51.

[0853] In one preferred embodiment, guide RNA Z3 (SEQ ID NO. 15) is used with guide RNA Z7 (SEQ ID NO. 19) and HDRT21 (SEQ ID NO. 52). In preferred embodiments, guide RNA Z3 (SEQ ID NO. 15) is used with guide RNA Z7 (SEQ ID NO. 19) and the homology arms of HDRT21 (SEQ ID NO. 52) without the knocked-in sequence of SEQ ID NO. 52.

[0854] In one preferred embodiment, guide RNA Z3 (SEQ ID NO. 15) is used with guide RNA Z8 (SEQ ID NO. 20) and HDRT21 (SEQ ID NO. 53). In preferred embodiments, guide RNA Z3 (SEQ ID NO. 15) is used with guide RNA Z8 (SEQ ID NO. 20) and the homology arms of HDRT21 (SEQ ID NO. 53) without the knocked-in sequence of SEQ ID NO. 53. In one preferred embodiment, guide RNA Z3 (SEQ ID NO. 15) is used with guide RNA Z5 (SEQ ID NO. 17) and HDRT22 (SEQ ID NO. 54). In preferred embodiments, guide RNA Z3 (SEQ ID NO. 15) is used with guide RNA Z5 (SEQ ID NO. 17) and the homology arms of HDRT22 (SEQ ID NO. 54) without the knocked-in sequence of SEQ ID NO. 54.

[0855] In one preferred embodiment, guide RNA Z3 (SEQ ID NO. 15) is used with guide RNA Z6 (SEQ ID NO. 18) and HDRT22 (SEQ ID NO. 55). In preferred embodiments, guide RNA Z3 (SEQ ID NO. 15) is used with guide RNA Z6 (SEQ ID NO. 18) and the homology arms of HDRT22 (SEQ ID NO. 55) without the knocked-in sequence of SEQ ID NO. 55.

[0856] In one preferred embodiment, guide RNA Z3 (SEQ ID NO. 15) is used with guide RNA Z7 (SEQ ID NO. 19) and HDRT23 (SEQ ID NO. 56). In preferred embodiments, guide RNA Z3 (SEQ ID NO. 15) is used with guide RNA Z7 (SEQ ID NO. 19) and the homology arms of HDRT23 (SEQ ID NO. 56) without the knocked-in sequence of SEQ ID NO. 56.

[0857] In one preferred embodiment, guide RNA Z3 (SEQ ID NO. 15) is used with guide RNA Z8 (SEQ ID NO. 20) and HDRT23 (SEQ ID NO. 57). In preferred embodiments, guide RNA Z3 (SEQ ID NO. 15) is used with guide RNA Z8 (SEQ ID NO. 20) and the homology arms of HDRT23 (SEQ ID NO. 57) without the knocked-in sequence of SEQ ID NO. 57.

[0858] In one preferred embodiment, guide RNA Z4 (SEQ ID NO. 16) is used with guide RNA Z5 (SEQ ID NO. 17) and HDRT24 (SEQ ID NO. 58). In preferred embodiments, guide RNA Z4 (SEQ ID NO. 16) is used with guide RNA Z5 (SEQ ID NO. 17) and the homology arms of HDRT24 (SEQ ID NO. 58) without the knocked-in sequence of SEQ ID NO. 58.

[0859] In one preferred embodiment, guide RNA Z4 (SEQ ID NO. 16) is used with guide RNA Z7 (SEQ ID NO. 19) and HDRT23 (SEQ ID NO. 59). In preferred embodiments, guide RNA Z4 (SEQ ID NO. 16) is used with guide RNA Z7 (SEQ ID NO. 19) and the homology arms of HDRT23 (SEQ ID NO. 59) without the knocked-in sequence of SEQ ID NO. 59.

[0860] In one preferred embodiment, guide RNA Z4 (SEQ ID NO. 16) is used with guide RNA Z8 (SEQ ID NO. 20) and HDRT23 (SEQ ID NO. 60). In preferred embodiments, guide RNA Z4 (SEQ ID NO. 16) is used with guide RNA Z8 (SEQ ID NO. 20) and the homology arms of HDRT23 (SEQ ID NO. 60) without the knocked-in sequence of SEQ ID NO. 60.

[0861] In one preferred embodiment, guide RNA B1 (SEQ ID NO. 21) is used with guide RNA B8 (SEQ ID NO. 28) and HDRT26 (SEQ ID NO. 61). In preferred embodiments, guide RNA B1 (SEQ ID NO. 21) is used with guide RNA B8 (SEQ ID NO. 28) and the homology arms of HDRT26 (SEQ ID NO. 61) without the knocked-in sequence of SEQ ID NO. 61.

[0862] In one preferred embodiment, guide RNA B2 (SEQ ID NO. 22) is used with guide RNA B6 (SEQ ID NO. 26) and HDRT27 (SEQ ID NO. 62). In preferred embodiments, guide RNA B2 (SEQ ID NO. 22) is used with guide RNA B6 (SEQ ID NO. 26) and the homology arms of HDRT27 (SEQ ID NO. 62) without the knocked-in sequence of SEQ ID NO. 62.

[0863] In one preferred embodiment, guide RNA B2 (SEQ ID NO. 22) is used with guide RNA B9 (SEQ ID NO. 29) and HDRT29 (SEQ ID NO. 63). In preferred embodiments, guide RNA B2 (SEQ ID NO. 22) is used with guide RNA B9 (SEQ ID NO. 29) and the homology arms of HDRT29 (SEQ ID NO. 63) without the knocked-in sequence of SEQ ID NO. 63. In one preferred embodiment, guide RNA B4 (SEQ ID NO. 24) is used with guide RNA B8 (SEQ ID NO. 28) and HDRT26 (SEQ ID NO. 64). In preferred embodiments, guide RNA B4 (SEQ ID NO. 24) is used with guide RNA B8 (SEQ ID NO. 28) and the homology arms of HDRT26 (SEQ ID NO. 64) without the knocked-in sequence of SEQ ID NO. 64.

[0864] In one preferred embodiment, guide RNA B5 (SEQ ID NO. 25) is used with guide RNA B8 (SEQ ID NO. 28) and HDRT26 (SEQ ID NO. 65). In preferred embodiments, guide RNA B5 (SEQ ID NO. 25) is used with guide RNA B8 (SEQ ID NO. 28) and the homology arms of HDRT26 (SEQ ID NO. 65) without the knocked-in sequence of SEQ ID NO. 65.

[0865] In one preferred embodiment, guide RNA T2 (SEQ ID NO. 2) is used with guide RNA T 11 (SEQ ID NO. 11) and HDRT10 (SEQ ID NO. 89). In one preferred embodiment, guide RNA T2 (SEQ ID NO. 2) is used with guide RNA T11 (SEQ ID NO. 11) and HDRT10 (SEQ ID NO. 89) without the knocked-in sequence of SEQ ID NO. 89.

[0866] In one preferred embodiment, guide RNA T2 (SEQ ID NO. 2) is used with guide RNA T12 (SEQ ID NO. 12) and HDRT10 (SEQ ID NO. 89). In one preferred embodiment, guide RNA T2 (SEQ ID NO. 2) is used with guide RNA T12 (SEQ ID NO. 12) and HDRT10 (SEQ ID NO. 89) without the knocked-in sequence of SEQ ID NO. 89.

[0867] In one preferred embodiment, guide RNA T2 (SEQ ID NO. 2) is used with guide RNA T7 (SEQ ID NO. 7) and HDRT10 (SEQ ID NO. 89). In one preferred embodiment, guide RNA T2 (SEQ ID NO. 2) is used with guide RNA T7 (SEQ ID NO. 7) and HDRT10 (SEQ ID NO. 89) without the knocked-in sequence of SEQ ID NO. 89.

[0868] In one preferred embodiment, guide RNA T3.1 (SEQ ID NO. 66) is used with guide RNA T10 (SEQ ID NO. 10) and HDRT11 (SEQ ID NO. 90). In one preferred embodiment, guide RNA T3.1 (SEQ ID NO. 66) is used with guide RNA T10 (SEQ ID NO. 10) and HDRT11 (SEQ ID NO. 90) without the knocked-in sequence of SEQ ID NO. 90.

[0869] In one preferred embodiment, guide RNA T3.1 (SEQ ID NO. 66) is used with guide RNA T8 (SEQ ID NO. 8) and HDRT12 (SEQ ID NO. 91). In one preferred embodiment, guide RNA T3.1 (SEQ ID NO. 66) is used with guide RNA T8 (SEQ ID NO. 8) and HDRT12 (SEQ ID NO. 91) without the knocked- in sequence of SEQ ID NO. 91.

[0870] In one preferred embodiment, guide RNA T3.1 (SEQ ID NO. 66) is used with guide RNA T9 (SEQ ID NO. 9) and HDRT18 (SEQ ID NO. 92). In one preferred embodiment, guide RNA T3.1 (SEQ ID NO. 66) is used with guide RNA T9 (SEQ ID NO. 9) and HDRT18 (SEQ ID NO. 92) without the knocked- in sequence of SEQ ID NO. 92.

[0871] In one preferred embodiment, guide RNA T3.1 (SEQ ID NO. 66) is used with guide RNA T 11 (SEQ ID NO. 11) and HDRT9 (SEQ ID NO. 47). In one preferred embodiment, guide RNA T3.1 (SEQ ID NO. 66) is used with guide RNA T11 (SEQ ID NO. 11) and HDRT9 (SEQ ID NO. 47) without the knocked-in sequence of SEQ ID NO. 47.

[0872] In one preferred embodiment, guide RNA T3.1 (SEQ ID NO. 66) is used with guide RNA T12 (SEQ ID NO. 12) and HDRT9 (SEQ ID NO. 47). In one preferred embodiment, guide RNA T3.1 (SEQ ID NO. 66) is used with guide RNA T12 (SEQ ID NO. 12) and HDRT9 (SEQ ID NO. 47) without the knocked-in sequence of SEQ ID NO. 47. In one preferred embodiment, guide RNA T3.1 (SEQ ID NO. 66) is used with guide RNA T7 (SEQ ID NO. 7) and HDRT9 (SEQ ID NO. 47). In one preferred embodiment, guide RNA T3.1 (SEQ ID NO.

[0873] 66) is used with guide RNA T7 (SEQ ID NO. 7) and HDRT9 (SEQ ID NO. 47) without the knocked-in sequence of SEQ ID NO. 47.

[0874] In one preferred embodiment, guide RNA T4.1 (SEQ ID NO. 67) is used with guide RNA T10 (SEQ ID NO. 10) and HDRT13 (SEQ ID NO. 93). In one preferred embodiment, guide RNA T4.1 (SEQ ID NO. 67) is used with guide RNA T10 (SEQ ID NO. 10) and HDRT13 (SEQ ID NO. 93) without the knocked-in sequence of SEQ ID NO. 93.

[0875] In one preferred embodiment, guide RNA T4.1 (SEQ ID NO. 67) is used with guide RNA T8 (SEQ ID NO. 8) and HDRT14 (SEQ ID NO. 94). In one preferred embodiment, guide RNA T4.1 (SEQ ID NO. 67) is used with guide RNA T8 (SEQ ID NO. 8) and HDRT14 (SEQ ID NO. 94) without the knocked- in sequence of SEQ ID NO. 94.

[0876] In one preferred embodiment, guide RNA T4.1 (SEQ ID NO. 67) is used with guide RNA T9 (SEQ ID NO. 9) and HDRT15 (SEQ ID NO. 95). In one preferred embodiment, guide RNA T4.1 (SEQ ID NO.

[0877] 67) is used with guide RNA T9 (SEQ ID NO. 9) and HDRT15 (SEQ ID NO. 95) without the knocked- in sequence of SEQ ID NO. 95. ln one preferred embodiment, guide RNA T4.1 (SEQ ID NO. 67) is used with guide RNA T 11 (SEQ ID NO. 11) and HDRT16 (SEQ ID NO. 96). In one preferred embodiment, guide RNA T4.1 (SEQ ID NO. 67) is used with guide RNA T11 (SEQ ID NO. 11) and HDRT16 (SEQ ID NO. 96) without the knocked-in sequence of SEQ ID NO. 96.

[0878] In one preferred embodiment, guide RNA T4.1 (SEQ ID NO. 67) is used with guide RNA T12 (SEQ ID NO. 12) and HDRT16 (SEQ ID NO. 96). In one preferred embodiment, guide RNA T4.1 (SEQ ID NO. 67) is used with guide RNA T12 (SEQ ID NO. 12) and HDRT16 (SEQ ID NO. 96) without the knocked-in sequence of SEQ ID NO. 96.

[0879] In one preferred embodiment, guide RNA T4.1 (SEQ ID NO. 67) is used with guide RNA T7 (SEQ ID NO. 7) and HDRT16 (SEQ ID NO. 96). In one preferred embodiment, guide RNA T4.1 (SEQ ID NO. 67) is used with guide RNA T7 (SEQ ID NO. 7) and HDRT16 (SEQ ID NO. 96) without the knocked- in sequence of SEQ ID NO. 96.

[0880] In one preferred embodiment, guide RNA T5.1 (SEQ ID NO. 68) is used with guide RNA T10 (SEQ ID NO. 10) and HDRT13 (SEQ ID NO. 93). In one preferred embodiment, guide RNA T5.1 (SEQ ID NO. 68) is used with guide RNA T10 (SEQ ID NO. 10) and HDRT13 (SEQ ID NO. 93) without the knocked-in sequence of SEQ ID NO. 93.

[0881] In one preferred embodiment, guide RNA T5.1 (SEQ ID NO. 68) is used with guide RNA T8 (SEQ ID NO. 8) and HDRT14 (SEQ ID NO. 94). In one preferred embodiment, guide RNA T5.1 (SEQ ID NO.

[0882] 68) is used with guide RNA T8 (SEQ ID NO. 8) and HDRT14 (SEQ ID NO. 94) without the knocked- in sequence of SEQ ID NO. 94.

[0883] In one preferred embodiment, guide RNA T5.1 (SEQ ID NO. 68) is used with guide RNA T9 (SEQ ID NO. 9) and HDRT15 (SEQ ID NO. 95). In one preferred embodiment, guide RNA T5.1 (SEQ ID NO. 68) is used with guide RNA T9 (SEQ ID NO. 9) and HDRT15 (SEQ ID NO. 95) without the knocked- in sequence of SEQ ID NO. 95. In one preferred embodiment, guide RNA T5.1 (SEQ ID NO. 68) is used with guide RNA T 11 (SEQ ID NO. 11) and HDRT16 (SEQ ID NO. 96). In one preferred embodiment, guide RNA T5.1 (SEQ ID NO. 68) is used with guide RNA T11 (SEQ ID NO. 11) and HDRT16 (SEQ ID NO. 96) without the knocked-in sequence of SEQ ID NO. 96.

[0884] In one preferred embodiment, guide RNA T5.1 (SEQ ID NO. 68) is used with guide RNA T12 (SEQ ID NO. 12) and HDRT16 (SEQ ID NO. 96). In one preferred embodiment, guide RNA T5.1 (SEQ ID NO. 68) is used with guide RNA T12 (SEQ ID NO. 12) and HDRT16 (SEQ ID NO. 96) without the knocked-in sequence of SEQ ID NO. 96.

[0885] In one preferred embodiment, guide RNA T5.1 (SEQ ID NO. 68) is used with guide RNA T7 (SEQ ID NO. 7) and HDRT16 (SEQ ID NO. 96). In one preferred embodiment, guide RNA T5.1 (SEQ ID NO. 68) is used with guide RNA T7 (SEQ ID NO. 7) and HDRT16 (SEQ ID NO. 96) without the knocked- in sequence of SEQ ID NO. 96. ln one preferred embodiment, guide RNA T6.1 (SEQ ID NO. 69) is used with guide RNA T10 (SEQ ID NO. 10) and HDRT13 (SEQ ID NO. 93). In one preferred embodiment, guide RNA T6.1 (SEQ ID NO. 69) is used with guide RNA T10 (SEQ ID NO. 10) and HDRT13 (SEQ ID NO. 93) without the knocked-in sequence of SEQ ID NO. 93.

[0886] In one preferred embodiment, guide RNA T6.1 (SEQ ID NO. 69) is used with guide RNA T8 (SEQ ID NO. 8) and HDRT14 (SEQ ID NO. 94). In one preferred embodiment, guide RNA T6.1 (SEQ ID NO. 69) is used with guide RNA T8 (SEQ ID NO. 8) and HDRT14 (SEQ ID NO. 94) without the knocked- in sequence of SEQ ID NO. 94.

[0887] In one preferred embodiment, guide RNA T6.1 (SEQ ID NO. 69) is used with guide RNA T9 (SEQ ID NO. 9) and HDRT15 (SEQ ID NO. 95). In one preferred embodiment, guide RNA T6.1 (SEQ ID NO. 69) is used with guide RNA T9 (SEQ ID NO. 9) and HDRT15 (SEQ ID NO. 95) without the knocked- in sequence of SEQ ID NO. 95.

[0888] In one preferred embodiment, guide RNA T6.1 (SEQ ID NO. 69) is used with guide RNA T 11 (SEQ ID NO. 11) and HDRT16 (SEQ ID NO. 96). In one preferred embodiment, guide RNA T6.1 (SEQ ID NO. 69) is used with guide RNA T11 (SEQ ID NO. 11) and HDRT16 (SEQ ID NO. 96) without the knocked-in sequence of SEQ ID NO. 96.

[0889] In one preferred embodiment, guide RNA T6.1 (SEQ ID NO. 69) is used with guide RNA T12 (SEQ ID NO. 12) and HDRT16 (SEQ ID NO. 96). In one preferred embodiment, guide RNA T6.1 (SEQ ID NO. 69) is used with guide RNA T12 (SEQ ID NO. 12) and HDRT16 (SEQ ID NO. 96) without the knocked-in sequence of SEQ ID NO. 96.

[0890] In one preferred embodiment, guide RNA T6.1 (SEQ ID NO. 69) is used with guide RNA T7 (SEQ ID NO. 7) and HDRT16 (SEQ ID NO. 96). In one preferred embodiment, guide RNA T6.1 (SEQ ID NO. 69) is used with guide RNA T7 (SEQ ID NO. 7) and HDRT16 (SEQ ID NO. 96) without the knocked- in sequence of SEQ ID NO. 96.

[0891] In one preferred embodiment, guide RNA Z1 (SEQ ID NO. 13) is used with guide RNA Z6 (SEQ ID NO. 18) and HDRT17 (SEQ ID NO. 49 of SEQ ID NO. 48). In one preferred embodiment, guide RNA Z1 (SEQ ID NO. 13) is used with guide RNA Z6 (SEQ ID NO. 18) and HDRT17 (SEQ ID NO. 49 or SEQ ID NO. 48) without the knocked-in sequence of SEQ ID NO. 49 or 48. In one preferred embodiment, guide RNA Z1 (SEQ ID NO. 13) is used with guide RNA Z5 (SEQ ID NO. 17) and HDRT17 (SEQ ID NO. 49 of SEQ ID NO. 48). In one preferred embodiment, guide RNA Z1 (SEQ ID NO. 13) is used with guide RNA Z5 (SEQ ID NO. 17) and HDRT17 (SEQ ID NO. 49 or SEQ ID NO. 48) without the knocked-in sequence of SEQ ID NO. 49 or 48.

[0892] In one preferred embodiment, guide RNA Z1 (SEQ ID NO. 13) is used with guide RNA Z8 (SEQ ID NO. 20) and HDRT27 (SEQ ID NO. 62). In one preferred embodiment, guide RNA Z1 (SEQ ID NO. 13) is used with guide RNA Z8 (SEQ ID NO. 20) and HDRT27 (SEQ ID NO. 62) without the knocked-in sequence of SEQ ID NO. 62.

[0893] In one preferred embodiment, guide RNA Z1 (SEQ ID NO. 13) is used with guide RNA Z7 (SEQ ID NO. 19) and HDRT23 (SEQ ID NO. 56, 57 or 60). In one preferred embodiment, guide RNA Z1 (SEQ ID NO. 13) is used with guide RNA Z7 (SEQ ID NO. 19) and HDRT23 (SEQ ID NO. 56, 57 or 60) without the knocked-in sequence of SEQ ID NO. 56,57 or 60.

[0894] In one preferred embodiment, guide RNA Z2 (SEQ ID NO. 14) is used with guide RNA Z8 (SEQ ID NO. 20) and HDRT23 (SEQ ID NO. 56, 57 or 60). In one preferred embodiment, guide RNA Z2 (SEQ ID NO. 14) is used with guide RNA Z8 (SEQ ID NO. 20) and HDRT23 (SEQ ID NO. 56, 57 or 60) without the knocked-in sequence of SEQ ID NO. 56,57 or 60.

[0895] In one preferred embodiment, guide RNA Z2 (SEQ ID NO. 14) is used with guide RNA Z7 (SEQ ID NO. 19) and HDRT23 (SEQ ID NO. 56, 57 or 60). In one preferred embodiment, guide RNA Z2 (SEQ ID NO. 14) is used with guide RNA Z7 (SEQ ID NO. 19) and HDRT23 (SEQ ID NO. 56, 57 or 60) without the knocked-in sequence of SEQ ID NO. 56,57 or 60.

[0896] In one preferred embodiment, guide RNA Z3 (SEQ ID NO. 15) is used with guide RNA Z6 (SEQ ID NO. 18) and HDRT19 (SEQ ID NO.97). In one preferred embodiment, guide RNA Z3 (SEQ ID NO. 15) is used with guide RNA Z6 (SEQ ID NO. 18) and HDRT19 (SEQ ID NO.97) without the knocked- in sequence of SEQ ID NO. 97.

[0897] In one preferred embodiment, guide RNA Z3 (SEQ ID NO. 15) is used with guide RNA Z5 (SEQ ID NO. 17) and HDRT19 (SEQ ID NO.97). In one preferred embodiment, guide RNA Z3 (SEQ ID NO. 15) is used with guide RNA Z5 (SEQ ID NO. 17) and HDRT19 (SEQ ID NO.97) without the knocked- in sequence of SEQ ID NO. 97. ln one preferred embodiment, guide RNA Z3 (SEQ ID NO. 15) is used with guide RNA Z8 (SEQ ID NO. 20) and HDRT25 (SEQ ID NO.98). In one preferred embodiment, guide RNA Z3 (SEQ ID NO.

[0898] 15) is used with guide RNA Z8 (SEQ ID NO. 20) and HDRT25 (SEQ ID NO.98) without the knocked- in sequence of SEQ ID NO. 98.

[0899] In one preferred embodiment, guide RNA Z3 (SEQ ID NO. 15) is used with guide RNA Z7 (SEQ ID NO. 19) and HDRT25 (SEQ ID NO.98). In one preferred embodiment, guide RNA Z3 (SEQ ID NO. 15) is used with guide RNA Z7 (SEQ ID NO. 19) and HDRT25 (SEQ ID NO.98) without the knocked- in sequence of SEQ ID NO. 98.

[0900] In one preferred embodiment, guide RNA Z4 (SEQ ID NO. 16) is used with guide RNA Z6 (SEQ ID NO. 18) and HDRT24 (SEQ ID NO.58). In one preferred embodiment, guide RNA Z4 (SEQ ID NO. 16) is used with guide RNA Z6 (SEQ ID NO. 18) and HDRT24 (SEQ ID NO.58) without the knocked- in sequence of SEQ ID NO. 98. In one preferred embodiment, guide RNA Z4 (SEQ ID NO. 16) is used with guide RNA Z8 (SEQ ID NO. 20) and HDRT28 (SEQ ID NO.99). In one preferred embodiment, guide RNA Z4 (SEQ ID NO. 16) is used with guide RNA Z8 (SEQ ID NO. 20) and HDRT28 (SEQ ID NO.99) without the knocked- in sequence of SEQ ID NO. 99.

[0901] In one preferred embodiment, guide RNA Z4 (SEQ ID NO. 16) is used with guide RNA Z7 (SEQ ID NO. 19) and HDRT28 (SEQ ID NO.99). In one preferred embodiment, guide RNA Z4 (SEQ ID NO. 16) is used with guide RNA Z7 (SEQ ID NO. 19) and HDRT28 (SEQ ID NO.99) without the knocked- in sequence of SEQ ID NO. 99.

[0902] In one preferred embodiment, guide RNA B5 (SEQ ID NO. 25) is used with guide RNA B6 (SEQ ID NO. 26) and HDRT29 (SEQ ID NO. 63). In one preferred embodiment, guide RNA B5 (SEQ ID NO. 25) is used with guide RNA B6 (SEQ ID NO. 26) and HDRT29 (SEQ ID NO. 63) without the knocked-in sequence of SEQ ID NO. 63.

[0903] In one preferred embodiment, guide RNA B5 (SEQ ID NO. 25) is used with guide RNA B9 (SEQ ID NO. 29) and HDRT29 (SEQ ID NO. 63). In one preferred embodiment, guide RNA B5 (SEQ ID NO. 25) is used with guide RNA B9 (SEQ ID NO. 29) and HDRT29 (SEQ ID NO. 63) without the knocked-in sequence of SEQ ID NO. 63.

[0904] In one preferred embodiment, guide RNA B5 (SEQ ID NO. 25) is used with guide RNA B7 (SEQ ID NO. 27) and HDRT30 (SEQ ID NO. 100). In one preferred embodiment, guide RNA B5 (SEQ ID NO. 25) is used with guide RNA B7 (SEQ ID NO. 27) and HDRT30 (SEQ ID NO. 100) without the knocked-in sequence of SEQ ID NO. 100.

[0905] In one preferred embodiment, guide RNA B1 (SEQ ID NO. 21) is used with guide RNA B6 (SEQ ID NO. 26) and HDRT31 (SEQ ID NO. 101). In one preferred embodiment, guide RNA B1 (SEQ ID NO. 21) is used with guide RNA B6 (SEQ ID NO. 26) and HDRT31 (SEQ ID NO. 101) without the knocked-in sequence of SEQ ID NO. 101.

[0906] In one preferred embodiment, guide RNA B1 (SEQ ID NO. 21) is used with guide RNA B9 (SEQ ID NO. 29) and HDRT31 (SEQ ID NO. 101). In one preferred embodiment, guide RNA B1 (SEQ ID NO. 21) is used with guide RNA B9 (SEQ ID NO. 29) and HDRT31 (SEQ ID NO. 101) without the knocked-in sequence of SEQ ID NO. 101.

[0907] In one preferred embodiment, guide RNA B1 (SEQ ID NO. 21) is used with guide RNA B7 (SEQ ID NO. 27) and HDRT31 (SEQ ID NO. 101). In one preferred embodiment, guide RNA B1 (SEQ ID NO. 21) is used with guide RNA B7 (SEQ ID NO. 27) and HDRT31 (SEQ ID NO. 101) without the knocked-in sequence of SEQ ID NO. 101.

[0908] In one preferred embodiment, guide RNA B4 (SEQ ID NO. 24) is used with guide RNA B6 (SEQ ID NO. 26) and HDRT29 (SEQ ID NO. 63). In one preferred embodiment, guide RNA B4 (SEQ ID NO. 24) is used with guide RNA B6 (SEQ ID NO. 26) and HDRT29 (SEQ ID NO. 63) without the knocked-in sequence of SEQ ID NO. 63.

[0909] In one preferred embodiment, guide RNA B4 (SEQ ID NO. 24) is used with guide RNA B9 (SEQ ID NO. 29) and HDRT29 (SEQ ID NO. 63). In one preferred embodiment, guide RNA B4 (SEQ ID NO. 24) is used with guide RNA B9 (SEQ ID NO. 29) and HDRT29 (SEQ ID NO. 63) without the knocked-in sequence of SEQ ID NO. 63. In one preferred embodiment, guide RNA B4 (SEQ ID NO. 24) is used with guide RNA B7 (SEQ ID NO. 27) and HDRT30 (SEQ ID NO. 100). In one preferred embodiment, guide RNA B4 (SEQ ID NO. 24) is used with guide RNA B7 (SEQ ID NO. 27) and HDRT30 (SEQ ID NO. 100) without the knocked-in sequence of SEQ ID NO. 100.

[0910] In one preferred embodiment, guide RNA B2 (SEQ ID NO. 22) is used with guide RNA B8 (SEQ ID NO. 28) and HDRT26 (SEQ ID NO.64). In one preferred embodiment, guide RNA B2 (SEQ ID NO. 22) is used with guide RNA B8 (SEQ ID NO. 28) and HDRT26 (SEQ ID NO.64) without the knocked- in sequence of SEQ ID NO. 64.

[0911] In one preferred embodiment, guide RNA B2 (SEQ ID NO. 22) is used with guide RNA B7 (SEQ ID NO. 27) and HDRT30 (SEQ ID NO. 100). In one preferred embodiment, guide RNA B2 (SEQ ID NO. 22) is used with guide RNA B7 (SEQ ID NO. 27) and HDRT30 (SEQ ID NO. 100) without the knocked-in sequence of SEQ ID NO. 100.

[0912] In one preferred embodiment, guide RNA E1 (SEQ ID NO. 70) is used with guide RNA E4 (SEQ ID NO. 73) and HDRT32 (SEQ ID NO. 102). In one preferred embodiment, guide RNA E1 (SEQ ID NO.

[0913] 70) is used with guide RNA E4 (SEQ ID NO. 73) and HDRT32 (SEQ ID NO. 102) without the knocked-in sequence of SEQ ID NO. 102.

[0914] In one preferred embodiment, guide RNA E1 (SEQ ID NO. 70) is used with guide RNA E5 (SEQ ID NO. 74) and HDRT32 (SEQ ID NO. 102). In one preferred embodiment, guide RNA E1 (SEQ ID NO. 70) is used with guide RNA E5 (SEQ ID NO. 74) and HDRT32 (SEQ ID NO. 102) without the knocked-in sequence of SEQ ID NO. 102.

[0915] In one preferred embodiment, guide RNA E2 (SEQ ID NO. 71) is used with guide RNA E4 (SEQ ID NO. 73) and HDRT32 (SEQ ID NO. 102). In one preferred embodiment, guide RNA E2 (SEQ ID NO.

[0916] 71) is used with guide RNA E4 (SEQ ID NO. 73) and HDRT32 (SEQ ID NO. 102) without the knocked-in sequence of SEQ ID NO. 102.

[0917] In one preferred embodiment, guide RNA E2 (SEQ ID NO. 71) is used with guide RNA E5 (SEQ ID NO. 74) and HDRT32 (SEQ ID NO. 102). In one preferred embodiment, guide RNA E2 (SEQ ID NO. 71) is used with guide RNA E5 (SEQ ID NO. 74) and HDRT32 (SEQ ID NO. 102) without the knocked-in sequence of SEQ ID NO. 102.

[0918] In one preferred embodiment, guide RNA E3 (SEQ ID NO. 72) is used with guide RNA E4 (SEQ ID NO. 73) and HDRT33 (SEQ ID NO. 103). In one preferred embodiment, guide RNA E3 (SEQ ID NO.

[0919] 72) is used with guide RNA E4 (SEQ ID NO. 73) and HDRT33 (SEQ ID NO. 103) without the knocked-in sequence of SEQ ID NO. 103.

[0920] A person skilled in the art is able to identify the homology arms of a HDR template, the gRNAs, and the knock-in sequence, based on the sequences provided in Table 1. In some embodiments, the combination of rRNA and HDRT lead to unexpectedly good results and high frequencies of the desired genetic modification. In some embodiments, the gRNAs work well due to their relative position in combination with the particular homology arms of the HDRT sequences, thus the homology arms of the HDRT sequences could be used without the specific knock-in sequence provided above, as described in the sequence table, but using some other construct to be knocked in, at essence in the same position of the target genome. In a further aspect of the invention, the invention relates to an isolated nucleic acid molecule, selected from the group consisting of: a) a nucleic acid molecule comprising a nucleotide sequence according to any one or more of SEQ ID NO 1-65, b) a nucleic acid molecule comprising the homology arms of an HDRT sequence presented in one or more of SEQ ID NO 38-65, although with any given knock-in sequence (different from sequences to be knocked-in of SEQ ID NO 38-65) c) a nucleic acid molecule which is complementary to a nucleotide sequence in accordance with a) or b); d) a nucleic acid molecule comprising a nucleotide sequence having sufficient sequence identity to be functionally analogous / equivalent to a nucleotide sequence according to a) through c), comprising preferably a sequence identity to a nucleotide sequence according to a) through c) of at least 70%, 80%, 90% or 95%; e) a nucleic acid molecule which, as a consequence of the genetic code, is degenerate to a nucleotide sequence according to a) through d); and / or f) a nucleic acid molecule according to a nucleotide sequence of a) through e) which is modified by deletions, additions, substitutions, translocations, inversions and / or insertions and is functionally analogous / equivalent to a nucleotide sequence according to a) through e).

[0921] In a further aspect of the invention, the invention relates to an isolated nucleic acid molecule, selected from the group consisting of: a) a nucleic acid molecule comprising a nucleotide sequence according to any one or more of SEQ ID NO 1-103, b) a nucleic acid molecule comprising the homology arms of an HDRT sequence presented in one or more of SEQ ID NO 38-65 or 89-103, although with any given knock-in sequence (different from sequences to be knocked-in of SEQ ID NO 38-65 or 89-103) c) a nucleic acid molecule which is complementary to a nucleotide sequence in accordance with a) or b); d) a nucleic acid molecule comprising a nucleotide sequence having sufficient sequence identity to be functionally analogous / equivalent to a nucleotide sequence according to a) through c), comprising preferably a sequence identity to a nucleotide sequence according to a) through c) of at least 70%, 80%, 90% or 95%; e) a nucleic acid molecule which, as a consequence of the genetic code, is degenerate to a nucleotide sequence according to a) through d); and / or f) a nucleic acid molecule according to a nucleotide sequence of a) through e) which is modified by deletions, additions, substitutions, translocations, inversions and / or insertions and is functionally analogous / equivalent to a nucleotide sequence according to a) through e).

[0922] In one embodiment the at least two guide RNAs (sgRNA) have a sequence selected from sequences SEQ ID No. 1 to 37 and 66 to 88.

[0923] In one embodiment the exogenous DNA repair template has a sequence selected from SEQ ID NO 38-65 and 89-103.

[0924] The term degenerate to (or degenerated into) refers to differences in a nucleotide sequence of a nucleic acid molecule, but according to the genetic code, do not lead to differences in amino acid sequence of the protein product of the nucleotide sequence after translation.

[0925] The various aspects of the invention are unified by, benefit from, are based on, and / or are linked by the structural and / or functional features, including functional properties and beneficial technical effects, of the in vitro method for modifying dsDNA in an eukaryotic cell described herein. The features disclosed in the context of the in vitro method also apply to and are considered disclosed in the context of the genetically modified eukaryotic cell and vice versa. Any features disclosed in further aspects of the invention, such as the kit, are also considered disclosed in the context of the in vitro method and vice versa.

[0926] DETAILED DESCRIPTION OF THE INVENTION

[0927] The present invention relates to an in vitro method for modifying double stranded DNA (dsDNA) in a eukaryotic cell.

[0928] Base editors (BEs) are a class of genome editing tools that enable the precise conversion of a single nucleotide in DNA or RNA without generating double-strand breaks (DSBs) or requiring donor DNA templates, making them a safer alternative to traditional gene-editing methods like CRISPR / Cas9, which rely on DSBs and can cause undesirable genomic damage. BEs typically consist of a catalytically impaired or dead Cas protein (e.g., Cas9 or Cast 2), which directs the system to a specific genomic locus via a guide RNA (gRNA) fused to a deaminase enzyme that chemically modifies target bases. This allows for the direct conversion of cytosine (C) to thymine (T) or adenine (A) to guanine (G), effectively enabling four types of base transitions (C to T, G to A, A to G, and T to C). Non-limiting examples for deaminase enzymes are adenosine deaminases (ADA) or cytidine deaminase (CDA). Usually, ADA is used for C-to-T base editing. CDA is used for A-to-G base editing. The skilled person in the art is able to select a corresponding deaminase enzyme without undue effort.

[0929] There are two main types of DNA base editors: cytosine base editors (CBEs) and adenine base editors (ABEs). CBEs convert a C*G base pair into a T*A base pair through the deamination of cytosine to uracil, which is recognized as thymine during DNA replication or repair. ABEs, on the other hand, convert an A*T base pair into a G*C base pair by deaminating adenine to inosine, which is read as guanine by the cell’s replication machinery. Both CBEs and ABEs use a single-stranded DNA bubble formed by the CRISPR-Cas complex as a substrate for deamination, ensuring high specificity to the target site. Glycosylase base editors (GBEs) are a specialized class of base editors that use a combination of Cas9 nickase, cytidine deaminase, and uracil-DNA glycosylase (Ung) to achieve base transversions, such as converting cytosine (C) to guanine (G). Unlike traditional base editors, which only induce base transitions, GBEs induce transversions by creating an apurinic / apyrimidinic (AP) site following the removal of uracil, the deaminated form of cytosine. This triggers the DNA repair machinery to complete the conversion. GBEs can perform C-to-G edits with high specificity and efficiency in both prokaryotic and mammalian cells, expanding the scope of possible genetic modifications beyond the traditional adenine-to-guanine and cytosine-to-thymine transitions. Modifications such as the integration of engineered variants of enzymes like APOBEC1 or the use of Saccharomyces cerevisiae Ung1 have enhanced GBE editing efficiency. These systems can achieve C-to-G conversions with varying efficiencies, offering the potential for addressing disease-causing mutations. Further improvements, such as the development of GBE2.0, have enhanced both the efficiency and purity of these edits. GBEs complement existing base editing tools by allowing for the correction of G / C-related mutations.

[0930] Within first-generation cytosine base editors (BE1), a catalytically impaired dCas9 protein is usually linked to a deaminase. The first generation of base editors encountered low base editing efficiency. CBEs mediate the deamination of cytosines, generating a uracil (U) base recognized as thymine by the cell’s DNA polymerases. However, uracil can be recognized and eliminated by the enzyme uracil DNA N-glycosylase (UNG) during the initiation of BER, resulting in fewer C to T edits and increased C to A or C to G edits. Second-generation cytosine base editors (BE2) additionally comprise a uracil glycosylase inhibitor (UGI) fused to the base editor. UGI inhibits the action of UNG, thus conserving the uracil base and achieving higher editing efficiency and purity. In third-generation base editors (BE3), instead of a catalytically dead Gas 9 (dCas9) protein, the nickase Cas9n is used, introducing a nick on the non-modified DNA strand. This nick serves to bias the cellular repair mechanisms to preferentially replace this strand and use the mutagenic intermediate as a template for repair, thus increasing editing efficiency. Fourth-generation cytosine base editors (BE4) comprise an additional copy of UGI, further reducing C-to-G or C-to-A conversions due to UNG.

[0931] Any base editor known in the art may be employed in accordance with the present invention.

[0932] As used herein, a “base editor system” refers to a system comprising a base editor and at least two guide RNA molecules. In embodiments, the base editor comprises an RNA-guided DNA nickase linked to a single-stranded DNA nucleobase-modifying enzyme or a nucleic acid encoding said base editor. In further embodiments, the guide RNA molecules comprise a first sgRNA that hybridizes to a second target sequence in an opposing strand of the dsDNA.

[0933] As used herein, an “editing window” refers to the precise range of DNA nucleotides within which the base editor is designed to target and modify. The editing window is determined by the gRNA sequence and the deaminase enzyme used in the base editor. The gRNA binds to a specific DNA sequence, and the deaminase enzyme acts on the nucleotides within a certain distance of the gRNA binding site. This distance defines the editing window. Depending on the base editor the base editing window is 3 to 15 nucleotides in length, such as 3, 4, 5, 6, 7, 8, 9, 10, 11 ,12, 13, 14 or 15 nucleotides in length. The term “double-stranded DNA (dsDNA)” molecule relates to two deoxyribonucleic acid polynucleotide strands that are bound together or hybridized by pairing the bases or nucleotides of the two strands through hydrogen bonds, resulting in double-stranded DNA. The method of the present invention can be performed using any dsDNA molecule, including, without limitation, genomic dsDNA of any origin or organism, chromosomal DNA, synthetic dsDNA, and amplified or isolated dsDNA. In the context of the present invention, the term “modifying” a double-stranded DNA refers to any kind of alteration, modification, or change of a dsDNA molecule. In particular, modifying relates to deleting, inserting, replacing, substituting, or translocating one or more nucleotides or pairs of nucleotides or nucleotide sequences from a dsDNA molecule. In the context of the invention, a dsDNA molecule to be modified is the dsDNA molecule, on which one or more of these modifications are introduced by the method of the invention.

[0934] As used herein, the term "hybridizing" refers to a reaction in which one or more polynucleotides react to form a complex that is stabilized via hydrogen bonding between the bases of the nucleotide residues. The hydrogen bonding may occur by Watson-Crick base pairing, Hoogstein binding, or any other sequence-specific manner. The complex may comprise two strands forming a duplex structure, three or more strands forming a multi-stranded complex, a single self-hybridizing strand, or any combination. A hybridization reaction may constitute a step in a more extensive process, such as the initiation of PCR or the cleavage of a polynucleotide by an enzyme. A sequence capable of hybridizing with a given sequence may be referred to as the "complement" of the given sequence, even if sequence complementarity is only partial.

[0935] Stringent conditions for hybridization refer to conditions under which a nucleic acid having complementarity to a target sequence hybridizes predominantly with the target sequence and substantially does not hybridize with non-target sequences. Stringent conditions are generally sequence-dependent and vary depending on several factors. In general, the longer the sequence, the higher the temperature at which it specifically hybridizes with its target sequence. Non-limiting examples of stringent conditions are described in detail in Tijssen (1993), Laboratory Techniques In Biochemistry And Molecular Biology-Hybridization With Nucleic Acid Probes Part I, Second Chapter "Overview of principles of hybridization and the strategy of nucleic acid probe assay", Elsevier, N.Y. When reference is made to a polynucleotide sequence, complementary or partially complementary sequences are also envisaged. These are preferably capable of hybridizing to the reference sequence under highly stringent conditions. Generally, to maximize the hybridization rate, relatively low-stringency hybridization conditions are selected: about 20 to 25° C. lower than the thermal melting point (Tm). The Tm is the temperature at which 50% of the specific target sequence hybridizes to a perfectly complementary probe in solution at a defined ionic strength and pH.

[0936] Generally, to require at least about 85% nucleotide complementarity of hybridized sequences, highly stringent washing conditions are selected to be about 5 to 15° C. lower than the Tm. In order to require at least about 70% nucleotide complementarity of hybridized sequences, moderately stringent washing conditions are selected to be about 15 to 30° C. lower than the Tm. Highly permissive (very low stringency) washing conditions may be as low as 50° C. below the Tm, allowing a high level of mismatching between hybridized sequences. Those skilled in the art will recognize that other physical and chemical parameters in the hybridization and wash stages can also be altered to affect the outcome of a detectable hybridization signal from a specific level of homology between target and probe sequences. Preferred highly stringent conditions comprise incubation in 50% formamide, 5xSSC, and 1% SDS at 42° C., or incubation in 5xSSC and 1 % SDS at 65° C., with a wash in 0.2xSSC and 0.1 % SDS at 65° C.

[0937] The term "insertion," per the present invention, is defined by the pertinent art and refers to incorporating one or more nucleotides into a nucleic acid molecule. Insertion of parts of genes, such as parts of exons or introns, as well as insertion of entire genes, is also encompassed by the term "insertion." When the number of inserted nucleotides is not dividable by three, the insertion can result in a frameshift mutation within a gene's coding sequence. Such frameshift mutations will alter the amino acids encoded by a gene following the mutation. In some cases, such a mutation will cause the active translation of the gene to encounter a premature stop codon, resulting in an end to translation and the production of a truncated protein. When the number of inserted nucleotides is instead dividable by three, the resulting insertion is an "in-frame insertion." In this case, the reading frame remains intact after the insertion, and translation will most likely run to completion if the inserted nucleotides do not code for a stop codon. However, because of the inserted nucleotides, the finished protein will contain, depending on the size of the insertion, one or multiple new amino acids that may affect the function of the protein.

[0938] The term "deletion," as used per the present invention, is defined by the pertinent art and refers to the loss of nucleotides or larger parts of genes, such as exons or introns, as well as entire genes. As defined with regard to the term "insertion," the deletion of a number of nucleotides that is not evenly dividable by three, in particular in a coding sequence, will lead to a frameshift mutation, causing all of the codons occurring after the deletion to be read incorrectly during translation, potentially producing a severely altered and most likely nonfunctional protein. If a deletion does not result in a frameshift mutation, i.e., because the number of nucleotides deleted is dividable by three, the resulting protein is nonetheless altered as the finished protein will lack. Depending on the size of the deletion, one or several amino acids may affect the function of the protein.

[0939] A “cell” in the sense of the present invention refers to, without limitation, any biological cell, which might be derived from any kind of organism, comprising unicellular organisms as well as multicellular organisms, such as any kind of plant or animal, including mammals, fish, amphibians, reptiles, birds, mollusks, arthropods, annelids, nematodes, flatworms, cnidarians, ctenophores and sponges. Cells of the present invention further comprise blood cells, stem cells, hematopoietic stem cells, hematopoietic stem and progenitor cells (HSPC), immune cells (such as B-cells, dendritic cells, granulocytes, innate lymphoid cells (ILCs), megakaryocytes, monocytes, macrophages, myeloid- derived Suppressor Cells (MDSC), natural killer (NK) cells, platelets, red blood cells (RBCs), T-cells or thymocytes), cancer cells, tumor cells and circulating tumor cells. In some embodiments of the invention, the cell in which the dsDNA is to be modified is a vertebrate cell, more preferably a mammalian cell, such as a human cell.

[0940] The term "introducing into the cell," as used herein, relates to any known method of introducing a protein or a nucleic acid molecule into a cell. Non-limiting examples include microinjection, infection with viral vectors, electroporation, and transfection, such as transfection using formulations with cationic lipids. The skilled person knows suitable methods for introducing the components of the present invention into a cell. Clustered Regularly Interspaced Short Palindromic Repeats (CRISPR) are a family of DNA sequences found in bacteria containing fragments of viral DNA from past infections. These fragments enable the bacteria to recognize and neutralize future attacks by similar viruses. Integral to the bacterial immune defense, CRISPR sequences have given rise to the CRISPR / Cas technology, which allows for precise and efficient gene editing in living organisms.

[0941] Sequences within CRISPR loci are transcribed and processed into CRISPR RNAs (crRNAs), which, in conjunction with trans-activating crRNAs (tracrRNAs), form complexes with CRISPR-associated (Cas) proteins. These complexes guide Cas nucleases to specific DNA targets through Watson- Crick base pairing between nucleic acids (Wiedenheft, B et al. (2012). Nature 482: 331-338; Horvath, P et al (2010). Science 327: 167-170; Fineran, PC et a. (2012). Virology 434: 202-209). The type II CRISPR nuclease system requires three key components: the Cas9 protein, mature crRNA, and tracrRNA. However, fusing the crRNA and tracrRNA into a single guide RNA (sgRNA) can simplify this system into two components. Re-targeting the Cas9 / sgRNA complex to new DNA sites is achievable by modifying a short sequence within the gRNA (Garneau, JE et al. (2010). Nature 468: 67-71 ; Deltcheva, E et al. (2011 ). Nature 471 : 602-607, Jinek, M et al. (2012) Science 337: 816-821).

[0942] CRISPR-Cas systems are RNA-guided adaptive immune systems of bacteria and archaea that provide sequence-specific resistance against viruses or other invading genetic material. This immune-like response is divided into two classes based on the architecture of the effector module responsible for target recognition and the cleavage of the invading nucleic acid (Makarova KS et al. Nat Rev Microbiol. 2015 Nov; 13(11):722-36.). Class 1 comprises multi-subunit Cas protein effectors, and Class 2 consists of a single large effector protein. Class 1 and 2 use CRISPR RNAs (crRNAs) to guide a Cas nuclease component to its target site, where it cleaves the invading nucleic acids. Due to their simplicity, Class 2 CRISPR-Cas systems are the most studied and widely applied for genome editing. The most widely used CRISPR-Cas system is CRISPR-Cas9.

[0943] It was demonstrated that the CRISPR / Cas9 system could be engineered for efficient genetic modification in mammalian cells. The only sequence limitation of the CRISPR / Cas system appears to derive from the necessity of a protospacer-adjacent motif (PAM) located immediately 3’ to the target site. The PAM sequence is specific to the species of Cas9. For example, the PAM sequence 5’-NGG-3’ is necessary to bind and cleavage DNA by the commonly used Cas9 from Streptococcus pyogenes. However, Cas9 variants with novel PAMs have been and may be engineered by directed evolution, thus expanding the number of potential target sequences. Cas9 complexed with the crRNA and tracrRNA undergoes a conformational change and associates with PAM motifs throughout the genome, interrogating the sequence directly upstream to determine sequence complementarity with the gRNA. The formation of a DNA-RNA heteroduplex at a matched target site allows for cleavage of the target DNA by the Cas9-RNA complex. These methods and mechanisms are well-known in the art. Upon binding of the Cas9-RNA complex (preferably comprising a sgRNA) to the target site / target sequence of a dsDNA to be modified, an RNA / DNA Cas9 complex is formed.

[0944] As known in the art, CRISPR / Cas9 has been exploited to develop potent tools for genome manipulation in animals, plants, and microorganisms. The RNA-guided Cas9 endonuclease first recognizes a 2- to 4-base-pair conserved sequence named the protospacer-adjacent motif (PAM), which flanks a target DNA site (target sequence). Upon binding to the PAM, Cas9 interrogates the flanking DNA sequences for base-pairing complementarity to a guide RNA. If the first 12 base pairs are complementary (the ‘seed’ sequence) of the guide RNA and the target DNA strand, RNA strand invasion accompanies local DNA unwinding to form an R-loop. Precise cleavage of each DNA strand by the RuvC and HNH domains of Cas9 generates a blunt double-strand DNA (dsDNA) break (DSB) at a position three base pairs upstream of the 3' edge of the protospacer sequence, measuring from the PAM.

[0945] CRISPR / Cas9 genome-editing experiments have been exploiting the host cell machinery to repair the genome precisely at the site of the Cas9-generated DSB. Mutations can arise either by non- homologous end joining (NHEJ) or homology-directed repair (HDR) of DSBs. NHEJ can produce small insertions or deletions (INDELs) at the cleavage site, whereas HDR uses a native (or engineered) DNA template to replace the targeted allele with an alternative sequence by recombination. Additional DNA repair pathways such as single-strand annealing, alternative end joining, microhomology-mediated joining, mismatch, and base- and nucleotide-excision repair can also produce genome edits.

[0946] As used herein, the term "INDEL" relates to the insertion or deletion of bases in an organism's genome generated upon repair after a dsDNA break.

[0947] Cas9, also named Csn1 , is a large protein that participates in crRNA biogenesis and the destruction of invading DNA. Cas9 has been described in different bacterial species such as S. thermophilus (Sapranauskas, Gasiunas, et al. 2011), listeria innocua (Gasiunas, Barrangou, et al. 2012; Jinek, Chylinski, et al. 2012) and S. pyogenes (Deltcheva, Chylinski, et al. 2011). The large Cas9 protein (>1200 amino acids) contains two predicted nuclease domains, namely the HNH (McrA-like) nuclease domain that is located in the middle of the protein and a split RuvC-like nuclease domain (RNase H fold) (Haft, Selengut, et al. 2005; Makarova, Grishin, et al. 2006). In wild-type Cas9, these two domains result in blunt cleavage of the invasive DNA within the same target sequence (protospacer) near the PAM (Jinek, Chylinski, et al. 2012).

[0948] Cas9 variants derived from Streptococcus pyogenes Cas9 (SpCas9) have been generated for use as nickases, dual nickases, or Fokl fusion variants. More recently, Cas9 orthologs and other nucleases derived from class 2 CRISPR-Cas systems, including Cpf1 and C2c1 , have been added to the CRISPR toolbox. These ongoing efforts to mine the abundant bacterial and archaeal CRISPR-Cas systems should increase the range of molecular tools available to researchers.

[0949] A nickase (or nicking enzyme or nicking endonuclease) is an enzyme that cuts one strand of a double-stranded DNA at a specific recognition nucleotide sequence known as a restriction site. Such enzymes hydrolyze (cut) only one strand of the DNA duplex to produce DNA molecules that are “nicked” rather than cleaved. Over 200 nicking enzymes have been studied, and 13 of these are available commercially and are routinely used for research and in commercial products.

[0950] Nickases can be and have been generated by mutating the nuclease domains of Cas9 independently of each other to create DNA nickase capable of introducing a single-strand cut with the same specificity as a regular CRISPR / Cas9 nuclease (Gasiunas G, Barrangou R, Horvath P, Siksnys V. Cas9-crRNA ribonucleoprotein complex mediates specific DNA cleavage for adaptive immunity in bacteria. Proc. Natl Acad. Sci. USA 109, E2579-E2586 (2012). The use of Cas9 nickases is essentially the same as the use of the fully functional enzyme. Cas9 nickase is created by mutating one of two Cas9 nuclease domains. Cas9 nickase creates a single-strand rather than a double-strand break. For example, the D10A mutation inactivates the RuvC domain, so this nickase cleaves only the target strand. Conversely, the H840 mutation in the HNH domain creates a non-target strand-cleaving nickase; instead of cutting both strands bluntly with WT Cas9 and one gRNA, a staggered cut using a Cas9 nickase and two gRNAs.

[0951] In other words, in RNA-guided DNA nickases, the naturally occurring endonuclease function of cleaving both strands of a double-stranded target DNA is altered into an endonuclease that cleaves (i.e., nicks) only one of the strands. Means and methods of modifying RNA-guided DNA endonuclease such as Cas9 are well known in the art and include, for example, the introduction of amino acid replacements into Cas9 that render one of the nuclease domains inactive. More specifically, aspartate can be replaced against alanine at position 10 of the Streptococcus pyogenes Cas9 (SpCas9 D10A; Cong et al. (2013) Science 339:819-823). Further examples are known in the art, for example the H840A replacement in SpCas9 (Mali P et al. Nat Biotechnol. 2013 Sep; 31 (9):833-8; Ran FA et al. Cell. 2013 Sep 12; 154(6): 1380-9).

[0952] The catalytic residues of the compact SpCas9 protein are those corresponding to amino acids D10, D31 , H840, H868, N882, and N891 or aligned positions using the CLUSTALW method on homologs of Cas Family members. Any of these residues can be replaced by any other amino acids, preferably by alanine residue. Mutation in the catalytic residues means either substituting with another amino acid or deleting or adding amino acids that induce the inactivation of at least one of the catalytic domains of Cas9. (Sapranauskas, Gasiunas et al. 2011 ; Jinek, Chylinski et al. 2012). In a particular embodiment, Cas9 may comprise one or several of the above mutations. In another particular embodiment, Cas9 may comprise only one of the two RuvC and HNH catalytic domains. In the present invention, Cas9 nickases of different species, Cas9 homologs, and Cas9 engineered, as well as functional variants thereof, can be used.

[0953] In the context of the present invention, the term “RNA-guided DNA endonuclease” refers to DNA endonucleases that interact with at least one RNA molecule. In the context of the present invention, the terms RNA-guided DNA endonuclease and RNA-guided endonuclease are used interchangeably. DNA endonucleases cleave the phosphodiester bond within a DNA polynucleotide chain. In the case of RNA-guided DNA endonuclease, the interacting RNA molecule may guide the RNA-guided DNA endonuclease to the site or location in a DNA where the endonuclease becomes active. In particular, the term RNA-guided DNA endonuclease refers to naturally occurring or genetically modified Cas nuclease components or CRISPR-Cas systems, which include, without limitation, multi-subunit Cas protein effectors of class 1 CRISPR-Cas systems as well as single large effector Cas proteins of class 2 systems. In the context of the present invention, DNA endonucleases functioning as nickases are used. Accordingly, in the context of the present invention, the RNA-guided DNA endonuclease is an RNA-guided DNA nickase. In the context of the invention, the RNA guiding the DNA nickase to the target site / sequence is a guide RNA.

[0954] Details of the technical application of CRISPR / Cas systems and suitable RNA-guided endonuclease are known to the skilled person and have been described in detail in the literature, for example, by Barrangou R et al. (Nat Biotechnol. 2016 Sep 8;34(9):933-941), Maeder ML et al. (Mol Then 2016 Mar;24(3):430-46) and Cebrian-Serrano A et al. (Mamm Genome. 2017; 28(7): 247-261). The present invention uses guide RNA-guided DNA nickases but is not limited to the use of a specific RNA-guided nickase and, therefore, comprises any given RNA-guided nickase in the sense of the present invention suitable for use in the method described herein.

[0955] Any RNA-guided DNA nickase known in the art may be employed in accordance with the present invention. A suitable nickase may be constructed based on an RNA-guided DNA endonuclease selected from the group comprising, without limitation, Gas proteins of class 1 CRISPR-Cas systems, such as Cas3, Cas8a, Cas5, Cas8b, Cas8c, CasWd, Cse1 , Cse2, Csy1 , Csy2, Csy3, GSU0054, Cast 0, Csm2, Cmr5, Csx11 , Csx10, and Csf1 ; Cas proteins of class 2 CRISPR-Cas systems, such as Cas9, Csn2, Cas4, Cpf1 , C2c1 , C2c3, and C2c2; corresponding orthologous enzymes / CRISPR effectors from various bacterial and archeal species; engineered CRISPR effectors with for example novel PAM specificities, increased fidelity, such as SpCas9- HF1 / eSpCas9, or altered functions. Particularly preferred are nickases based on RNA-guided DNA endonuclease selected from the group comprising Streptococcus pyogenes Cas9 (SpCas9), Staphylococcus aureus Cas9, Streptococcus thermophilus Cas9, Neisseria meningitidis Cas9 (NmCas9), Francisella novicida Cas9 (FnCas9), Campylobacter jejuni Cas9 (CjCas9), Cast 2a (Cpf1) and Cast 3a (C2C2) (Makarova KS et al. (November 2015). Nature Reviews Microbiology. 13 (11): 722-36).

[0956] In accordance with the method of the invention, the RNA-guided DNA nickase may be introduced as a protein, but alternatively, it may also be introduced as a nucleic acid molecule encoding said protein. It will be appreciated that the nucleic acid molecule encodes said RNA-guided DNA nickase in an inexpressible form such that expression in the cell results in a functional RNA-guided DNA nickase protein. Means and methods to ensure the expression of a functional polypeptide are well- known in the art.

[0957] For example, the coding sequences for the nickase may be comprised in a vector, such as a plasmid, cosmid, virus, bacteriophage, or another vector used conventionally, e.g., in genetic engineering. The coding sequences inserted in the vector can, e.g., be synthesized by standard methods or isolated from natural sources. The coding sequences may further be ligated to transcriptional regulatory elements and / or to other amino acid encoding sequences. Such regulatory sequences are well known to those skilled in the art and include, without being limiting, regulatory sequences ensuring the initiation of transcription, internal ribosomal entry sites (IRES), and optionally regulatory elements ensuring termination of transcription and stabilization of the transcript. Non-limiting examples for regulatory elements ensuring the initiation of transcription comprise a translation initiation codon, transcriptional enhancers such as e.g. the SV40-enhancer, insulators and / or promoters, such as the cytomegalovirus (CMV) promoter, SV40-promoter, RSV-promoter (Rous sarcome virus), the lacZ promoter, chicken beta-actin promoter, CAG-promoter (a combination of chicken beta-actin promoter and cytomegalovirus immediate-early enhancer), the gai10 promoter, human elongation factor 1a-promoter, A0X1 promoter, GAL1 promoter CaM-kinase promoter, the lac, trp or tac promoter, the lacUVS promoter, the autographa californica multiple nuclear polyhedrosis virus (AcMNPV) polyhedral promoter or a globin intron in mammalian and other animal cells. Non-limiting examples for regulatory elements ensuring transcription termination include the V40-poly-A site, the tk-poly-A site, or the SV40, lacZ, or AcMNPV polyhedral polyadenylation signals, which are to be included downstream of the nucleic acid sequence of the invention. Additional regulatory elements may include translational enhancers, Kozak sequences, and intervening sequences flanked by donor and acceptor sites for RNA splicing. Moreover, elements such as the origin of replication, drug resistance genes, or regulators (as part of an inducible promoter) may also be included.

[0958] As used herein, "nucleic acid" shall mean any nucleic acid molecule, including, without limitation, DNA, RNA, and hybrids or modified variants thereof. An "exogenous nucleic acid" or "exogenous genetic element" relates to any nucleic acid introduced into the cell that is not a component of the cell's "original" or "natural" genome. Exogenous nucleic acids may be integrated or nonintegrated in the genetic material of the target cell or relate to stably transduced nucleic acids.

[0959] Nucleic acid molecules encoding said RNA-guided DNA nickase include DNA, such as cDNA or genomic DNA, as well as RNA, particularly mRNA. The skilled person will readily appreciate that more than one nucleic acid molecule may encode an RNA-guided DNA nickase under the present invention due to the degeneracy of the genetic code. Degeneracy results because a triplet code designates 20 amino acids and a stop codon. Because four bases are utilized to encode genetic information, triplet codons are required to produce at least 21 different codes. The possible e possibilities for bases in triplets give 64 possible codons, meaning that some degeneracy must exist. As a result, some amino acids are encoded by more than one triplet, i.e. , by up to six. The degeneracy mostly arises from alterations in the third position in a triplet. This means that nucleic acid molecules having different sequences but still encoding the same RNA-guided DNA endonuclease can be employed by the present invention.

[0960] The nucleic acid molecules in the present invention may be of natural and / or (semi) synthetic origin. Thus, they may, for example, be nucleic acid molecules that have been synthesized according to conventional protocols of organic chemistry. The person skilled in the art is familiar with the preparation and use of said probes (see, e.g., Sambrook and Russel, "Molecular Cloning, A Laboratory Manual," Cold Spring Harbor Laboratory, N.Y. (2001)).

[0961] In embodiments, the nucleic acids encoding for the RNA-guided endonuclease, base editor, or prime editor, in the context of the present invention, includes nucleic acids, preferably mRNA, containing modified backbones, non-natural internucleoside linkages, modified nucleotides, poly-(A)-tail and / or untranslated regions (UTRs). The invention encompasses nucleic acids with at least one structural modification that provides improved stability, reduced immunogenicity, and / or half-life of said nucleic acid molecule post-administration in a cell and / or organism compared to a structurally unmodified nucleic acid of the same sequence.

[0962] Non-limiting examples of modified oligonucleotide backbones include, for example, phosphorothioates, chiral phosphorothioates, phosphorodithioates, phosphotriesters, aminoalkylphosphotriesters, methyl, and other alkyl phosphonates. Various salts, mixed salts, and free acid forms are also included.

[0963] The nucleic acid molecules used in accordance with the invention may be nucleic acid mimicking molecules known in the art, such as synthetic or semi-synthetic derivatives of nucleic acid molecules and mixed polymers. They may contain additional non-natural or derivatized nucleotide bases, as will be readily appreciated by those skilled in the art. Nucleic acid mimicking molecules or nucleic acid derivatives, according to the invention, include without being limiting, phosphorothioate nucleic acid, phosphoramidite nucleic acid, morpholino nucleic acid, hexitol nucleic acid (HNA), peptide nucleic acid (PNA) and locked nucleic acid (LNA).

[0964] Modified nucleobases include other synthetic and natural nucleobases such as N’- methylpseudouridine, 5-methylcytosine (5-me-C), 5-hydroxymethyl cytosine, xanthine, hypoxanthine, 2-aminoadenine, 6-methyl and other alkyl derivatives of adenine and guanine, 2-propyl and other alkyl derivatives of adenine and guanine, 2 -thiouracil, 2 -thiothymine and 2-thiocytosine, 5-halouracil and cytosine, 5-propynyl (-CEC-CH3) uracil and cytosine and other alkynyl derivatives of pyrimidine bases, 6-azo uracil, cytosine and thymine, 5-uracil (pseudouracil), 4-thiouracil, 8-halo, 8-amino, 8- thiol, 8-thioalkyl, 8-hydroxyl and other 8-substituted adenines and guanines, 5-halo particularly 5- bromo, 5-trifluoromethyl and other 5-substituted uracils and cytosines, 7-methylguanine and 7- methyladenine, 2-F-adenine, 2-amino-adenine, 8-azaguanine and 8-azaadenine, 7-deazaguanine and 7-deazaadenine and 3-deazaguanine and 3-deazaadenine. Further modified nucleobases include tricyclic pyrimidines such as phenoxazine cytidine(1 H-pyrimido[5,4-b][1 ,4]benzoxazin-2(3H)- one), phenothiazine cytidine (1 H-pyrimido[5,4-b][1 ,4]benzothiazin-2(3H)-one), G-clamps such as a substituted phenoxazine cytidine (e.g. 9-(2-aminoethoxy)-H-pyrimido[5,4-b][1 ,4]benzoxazin-2(3H)- one), carbazole cytidine (2H-pyrimido[4,5-b]indol-2-one), pyridoindole cytidine (H- pyrido[3',2':4,5]pyrrolo[2,3-d]pyrimidin-2-one). Modified nucleobases may also include those in which the purine or pyrimidine base is replaced with other heterocycles, for example 7-deaza-adenine, 7- deazaguanosine, 2-aminopyridine and 2-pyridone.

[0965] It is not necessary for all positions in a given nucleic acid to be uniformly modified, and in fact, more than one of the aforementioned modifications may be incorporated in a single compound or even at a single nucleoside within a nucleic acid.

[0966] A skilled person is aware of appropriate methods of nucleic acid synthesis. Modern techniques enable rapid and inexpensive custom-made nucleic acids of a desired sequence. For example, a common process relates to solid-phase synthesis using the phosphoramidite method and phosphoramidite building blocks derived from protected 2'-deoxynucleosides (dA, dC, dG, and T), ribonucleosides (A, C, G, and U), or chemically modified nucleosides, e.g., LNA, BNA. To obtain the desired oligonucleotide, the building blocks (modified or naturally occurring) are sequentially coupled to the growing oligonucleotide chain in the order required by the sequence of the product. Upon the completion of the chain assembly, the product is released from the solid phase to the solution, deprotected, and collected. Products may be isolated by high-performance liquid chromatography (HPLC) to obtain the desired nucleic acids in high purity if required.

[0967] Furthermore, the present invention's method comprises introducing at least two guide RNAs into the cell. In the context of the present invention, a "guide RNA" refers to RNA molecules interacting with RNA-guided DNA endonuclease, such as a nickase, leading to the recognition of the target sequence to be cleaved by the RNA-guided DNA endonuclease. According to the present invention, the term "guide RNA" therefore comprises, without limitation, target sequence-specific CRISPR RNAs (crRNA), trans-activating crRNAs (tracrRNA), and chimeric single guide RNAs (sgRNA).

[0968] Guide RNA (gRNA) and single guide RNA (sgRNA) are components of the CRISPR-Cas system, primarily used for gene editing. gRNA is a short RNA sequence that guides Cas9 endonuclease or other Gas proteins to specific DNA sites for targeted cleavage. In bacteria and archaea, gRNAs are integral to the adaptive immune response, directing Cas enzymes to degrade foreign DNA. sgRNA, a more streamlined version, combines the crRNA (which binds to the target DNA) and tracrRNA (which activates Cas9) into a single molecule, enhancing efficiency and ease of use in CRISPR applications. Both gRNA and sgRNA enable precise targeting of genes, facilitating various applications in genetic research and biotechnology. As used herein, the terms “single guide RNA” and “guide RNA” can be used interchangeably. crRNAs differ depending on the RNA-guided endonuclease and the CRISPR / Cas system but typically contain a target-specific sequence of between 20 to 72 nucleotides in length, flanked by two direct repeats (DR) of a length of between 21 to 46 nucleotides. In the case of S. pyogenes, the DRs are 36 nucleotides long, and the target sequence is 30 nucleotides long. The 3' located DR of the crRNA is complementary to and hybridizes with the corresponding tracr RNA, which in turn binds to the Cas9 protein.

[0969] As used herein, the term "trans-activating crRNA (tracrRNA)" refers to a small RNA that is complementary to and base pairs with a pre-crRNA (3' located DR of the crRNA), thereby forming an RNA duplex. This pre-crRNA is then cleaved by an RNA-specific ribonuclease to form a crRNA / tracrRNA hybrid, which subsequently acts as a guide for the endonuclease Cas9, which cleaves the invading nucleic acid.

[0970] As described herein, the genes encoding the elements of a CRISPR / Cas system, such as Cas9, tracrRNA, and crRNA, are typically organized in an operon(s). DR sequences functioning together with RNA-guided endonucleases, such as Cas9 proteins of other bacterial species, may be identified by bioinformatic analysis of sequence repeats occurring in the respective CRISPR / Cas operons and by experimental binding studies of Cas9 protein and tracrRNA together with putative DR sequence flanked target sequences.

[0971] Alternatively, a chimeric single guide RNA sequence comprising such a target sequence-specific crRNA and tracrRNA may be employed. Such a chimeric (ch) RNA may be designed by the fusion of a target-specific sequence of 20 or more nucleotides (nt) with a part or the entire DR sequence (defined as part of a crRNA) with the entire or part of a tracrRNA, as shown by (Jinek et al. Science 337:816-821). Within the chimeric RNA, a segment of the DR and the tracrRNA sequence are complementary and able to hybridize and form a hairpin structure. In the context of the invention, sgRNAs as guide RNAs are preferred.

[0972] Moreover, the two guide RNAs of the present invention differing in their target-specific sequence may also be encoded by a nucleic acid molecule, which is introduced into the cell. The definitions and preferred embodiments concerning the nucleic acid molecule encoding the nickase apply equally to the nucleic acid molecule encoding these RNAs. Regulatory elements for expressing RNAs are known to one skilled in the art, such as a U6 promoter.

[0973] The present invention relates to the generation of single-strand breaks of the dsDNA molecule to be modified, wherein the dsDNA molecule comprises at least two target sequences, which are targeted by at least two guide RNAs. The first guide RNA targets / hybridizes to a first target sequence in the dsDNA to be modified, and the second guide RNA hybridizes to a second target sequence in the dsDNA to be modified, preferably with the target-specific sequence of the guide RNA. In accordance with the present invention, a "target sequence" is a nucleotide sequence in the dsDNA molecule that is recognized by the guide RNA and is associated with the RNA-guided nickase due to the target-specific sequence comprised by the guide RNA. The target sequence is at least partially complementary to the target-specific sequence of the guide RNA and is associated with a so-called protospacer adjacent motif (PAM). PAM is a 2-6 base pair DNA sequence located adjacent to the target sequence and can be located either at the 5’-end (for example, for the Crispr / Cpf1 system) or at the 3’-end of the target sequence (for example, for the Crispr / Cas9 system), depending on the Crispr / Cas system employed. An RNA-guided nickase, such as a Cas9 or Cpf1 -based nickase, will not successfully bind to and cleave the targeted dsDNA molecule if the recognized target sequence is not associated with a PAM sequence. Forming a DNA-RNA heteroduplex between the target sequence and the target-specific sequence of the guide RNA with the bound nickase allows for cleavage of the target DNA by this guide RNA / DNA nickase complex. Cleavage of the targeted dsDNA molecule occurs within the target sequence or at a site in close proximity or adjacent to the target sequence, depending on the function of the used RNA-guided nickase and CRISPR / Cas system.

[0974] For example, in the case of SpCas9, the DNA target sequence of at least 20 nucleotides is located directly upstream / at the 5’-end of an invariant 5’-NGG-3' PAM. Correct pairing of the guide RNA to the DNA target sequence leads to the generation of a cut in the dsDNA molecule ("cleavage" of the dsDNA molecule by SpCas9) 3 base-pairs (bp) upstream of the PAM within the target sequence. However, SpCas9 mutants have been developed with altered PAM requirements (e.g. minimal PAMs, e.g. SpCas9-NG, SpyR).

[0975] In the case of the CRISPR / Cpf1 system, the Cpf1-crRNA complex cleaves target DNA by identification of a target sequence that may be located downstream / at the 3' end of a protospacer adjacent motif (for example, 5'-YTN-3' (where "Y" is a pyrimidine and "N" is any nucleobase) or 5'- TTN-3'). Cpf1 can introduce a sticky end / staggered end DNA double-strand breaks. In the case of AsCpfl and LbCpfl , a double-strand break with a 4 nucleotides overhang can be generated, which can occur 19 bp downstream of the PAM on the targeted (+)-strand and 23 bp downstream of the PAM on the (-)-strand. Corresponding nickases function analogously.

[0976] The exact site of the cut depends on the CRISPR / Cas system or the RNA-guided endonuclease / nickase employed in the method of the invention and can, therefore, be determined by the person skilled in the art upon selection of the RNA-guided DNA nickase.

[0977] In the context of the present invention, a cleavage site in close proximity to the target sequence is located within 100 nucleotides or base pairs upstream or downstream from the 5’- or 3’-end of the target sequence. Preferably, the double-strand break is generated within 90, 80, 70, 60, 50, 40, 30, 20, 10, 5, 4, 3, 2, or 1 nucleotides / base pairs upstream or downstream from the 5’- or 3’-end of the target sequence or within the target sequence.

[0978] In the context of the present invention, the target sequence may also be called a "protospacer." The term "target site" may refer to a location or sequence in the dsDNA molecule comprising the target sequence and an associated PAM.

[0979] In the context of the present invention, a "single strand break" (SSB) refers to a disruption involving only one of the two strands within a double-stranded DNA (dsDNA) molecule. Such a break does not result in the physical dissociation of the upstream and downstream segments of the dsDNA at the site of the lesion. In contrast, a "double-strand break" or "DSB" refers to the interruption of both strands of a dsDNA molecule, leading to the separation of the parts of the dsDNA molecule that lie upstream and downstream of the side of the double-strand break.

[0980] In the context of the present invention, at least two SSBs can occur due to the cleavage of both strands of a dsDNA, which is to be modified by an RNA-guided nickase. In embodiments, a DSB may occur due to cleavage (introduction of SSBs) by the nickase at the two target sites on the opposing strands of the dsDNA to be modified, i.e. the (+)- and the (-)-strand, despite the spacer / distance between the SSBs.

[0981] In the context of the invention, two SSBs are produced intentionally by RNA-guided nucleases to achieve modification of the dsDNA through homology-directed repair.

[0982] Cellular DNA repair mechanisms implicated in the repair of double-strand breaks (DSBs) include homology-directed repair (HDR), alternative non-homologous end joining (a-NHEJ), and classical non-homologous end joining (c-NHEJ). The c-NHEJ pathway remains active throughout all phases of the cell cycle. In contrast, HDR is confined to the S and G2 phases, requiring homologous DNA sequences to mediate repair. During mitosis, DSB repair is entirely suppressed to prevent telomere fusion, a critical safeguard for chromosomal integrity. In the G1 (and GO) phases, as well as in quiescent cells, c-NHEJ predominates due to the silencing of HDR. However, in the S and G2 phases, all repair pathways are active and complete. However, c-NHEJ often predominates when DSBs are induced in a population of cycling cells, resulting in a spectrum of edited alleles.

[0983] “Homology-directed repair (HDR)” is a precise cellular mechanism for repairing double-stranded DNA breaks. The most well-known form of HDR is homologous recombination, utilized primarily during the G2 and S phases when a homologous DNA template is present in the nucleus. HDR also encompasses other pathways, such as single-strand annealing and break-induced replication. In the absence of a homologous DNA sequence, the cell resorts to non-homologous end joining (NHEJ), which does not require sequence homology for repair.

[0984] In proliferating cells, homology-directed repair (HDR) is facilitated primarily through the homologous recombination (HR) pathway. This native mechanism utilizes intact homologous sequences from sister chromatids as a template to repair double-strand breaks (DSBs), leading to the restoration of the wild-type allele. The art has demonstrated that precise sequence modifications can be introduced at targeted DSBs by leveraging the HR pathway. Targeted modifications can be achieved by providing an exogenous DNA donor template — also referred to as a repair template — containing sequences homologous to the DSB ends. The sequence located between the homologous ends, whether it involves an insertion or a replacement, is seamlessly integrated into the target locus during HR, allowing for the generation of precisely modified ‘knock-in’ alleles. These modifications are particularly useful for codon replacements or the insertion of reporter genes. Notably, large sequence insertions often necessitate the use of double-stranded DNA donor templates, such as plasmid-based gene-targeting vectors with homology arms of at least 100 base pairs (preferably 500 base pairs or more), to ensure efficient integration.

[0985] The HDR-mediated repair of DSBs requires resection at the DSB ends, generating 3' singlestranded DNA (ssDNA) overhangs. These overhangs subsequently anneal with a homologous DNA sequence, which serves as a template for DNA repair synthesis across the DSB. HDR predominantly occurs via the high-fidelity homologous recombination repair (HRR) pathway but can also proceed through error-prone mechanisms such as single-strand annealing (SSA) or microhomology-mediated end joining (MMEJ). HRR and SSA share initial steps involving ATM signaling, formation of ionizing radiation-induced foci (IRIF), extensive resection of DSB ends, and activation of ATR signaling. In HRR, the 3’-ssDNA overhangs anneal with complementary sequences on sister chromatids, ensuring accurate repair. In contrast, SSA involves annealing the 3’-ssDNA overhangs via homologous direct repeats, leading to the deletion of one repeat and the intervening sequence during the repair process. While HRR and SSA rely on the annealing of long, highly homologous DNA sequences, MMEJ utilizes much shorter homology regions (typically up to 20 nucleotides), making it a more promiscuous repair mechanism prone to joining unrelated DNA fragments. The error-prone nature of MMEJ is further exacerbated by the involvement of the low- fidelity DNA polymerase theta (POLQ), which mediates DNA synthesis during this repair process. For comprehensive reviews on these mechanisms, refer to Khanna KK, Nat Genet 2001 ; Thompson LH and Schild D, Mutat Res 2001 ; Thompson LH and Schild D, Mutat Res 2002; Ciccia A and Elledge SJ, Mol. Cell 2010.

[0986] A DNA repair modulator is an agent, either naturally occurring or synthetic, that influences the efficiency or dynamics of DNA repair mechanisms. DNA repair modulators can target specific enzymes or proteins involved in DNA repair pathways, such as DNA glycosylases, AP endonucleases, DNA polymerases, and DNA ligases, which work sequentially to recognize, excise, and repair damaged DNA segments. Modulators may either enhance or inhibit the kinetics of DNA repair. Enhancers increase the activity of repair enzymes or facilitate the recruitment of repair complexes to damaged sites, while inhibitors may reduce or inhibit repair processes by interfering with enzymatic activity or repair intermediates.

[0987] In embodiments, the method of the present invention involves contacting a cell with one or more DNA-PK inhibitors or DNA PolQ inhibitors.

[0988] DNA-PK and DNA PolQ inhibitors are established inhibitors targeting DNA repair pathways. DNA-PK inhibitors are compounds that target DNA-PK, a protein kinase that plays a crucial role in NHEJ. By inhibiting DNA-PK, these compounds can interfere with the repair of DSBs, leading to cell death. Non-limiting examples of DNA-PK inhibitors are NU7026, NU7441 , M3814, AZD7648, and KU- 0060648.

[0989] DNA PolQ is a DNA polymerase that is also involved in NHEJ, but also considered to be important for microhomology-mediated end joining (MMEJ). It is particularly active in repairing DSBs that are difficult to repair by other mechanisms. Inhibitors of DNA PolQ can also interfere with NHEJ and lead to cell death.

[0990] Any of DNA-PK inhibitor or DNA PolQ inhibitor known in the art is suitable for the method of the present invention. A skilled person is capable of selecting a suitable inhibitor without undue effort.

[0991] In the context of the present invention, the terms “homology arms” or “homology arms that are targeted to the dsDNA molecule to be modified” refer to regions or sequences of the exogenous nucleic acid molecule that are homologous to the sequences in proximity to the two double-strand break ends of the dsDNA molecule to be modified. Homology arms may have 90 %, preferably 95 %, 97 %, 98 %, 99 %, or 100 % sequence identity to the corresponding sequences of the dsDNA molecule to be modified. Homology arms have sufficient sequence identity to ensure specific binding to the target sequence. Methods to evaluate the identity level between two nucleic acid sequences are well known in the art. For example, the sequences can be aligned electronically using suitable computer programs known in the art. Such programs comprise BLAST (Altschul et al. (1990) J. Mol. Biol. 215, 403), variants thereof such as WU-BLAST (Altschul and Gish (1996) Methods Enzymol. 266, 460), FASTA (Pearson and Lipman (1988) Proc. Natl. Acad. Sci. USA 85, 2444) or implementations of the Smith-Waterman algorithm (SSEARCH, Smith and Waterman (1981) J. Mol. Biol., 147, 195). These programs, in addition to providing a pairwise sequence alignment, also report the sequence identity level (usually in percent identity) and the probability for the occurrence of the alignment by chance (P-value). In accordance with the present invention, BLAST is preferred to be used to determine the level of identification between two nucleic acid sequences.

[0992] As used herein, the term “exogenous DNA repair template” refers to an exogenous DNA molecule that is introduced into the cell and that comprises a sequence to be integrated into the dsDNA molecule to be modified (the DNA template sequence). In the context of the present invention, the template sequence replaces a DNA sequence of the dsDNA molecule to be modified that is in proximity to, adjacent to, and / or between the two SSBs introduced into the dsDNA molecule.

[0993] In preferred embodiments, the sequence to be replaced is a mutated sequence resulting in a functionally impaired or inactive gene product encoded by the dsDNA molecule to be modified, and the DNA template sequence comprises a corrected / functional version of the respective sequence so that replacement of the mutated sequence by the template sequence leads to expression of a functional gene product by the modified dsDNA molecule.

[0994] In further embodiments, the DNA template sequence comprises an endogenous and / or exogenous sequence to the cell. Examples of a sequence to be integrated include DNA sequences encoding a protein, a non-coding RNA (e.g., a microRNA) or a landing pad sequence for a recombinase (e.g. att sites for bacteriophage derived large serine integrases). In embodiments, the template sequence can comprise a coding sequence that is operably linked to a regulatory sequence for controlling gene expression, such as promoter or promoter / enhancer sequence sequences. In embodiments, a reporter gene or reporter cassette can be integrated into the dsDNA to be modified by replacing a DNA sequence in proximity to, adjacent to, and / or between the two SSBs.

[0995] The homologous sequences required for recombination, commonly referred to as homology arms, are designed to flank the sequence to be replaced and are homologous to regions located upstream and downstream of the target site. It is understood that these homologous regions are part of the endogenous DNA sequence targeted for modification. In some embodiments, the upstream homology arm consists of a nucleic acid sequence that shares significant sequence identity with the genome sequence upstream of the integration or replacement site within the dsDNA molecule. Likewise, the downstream homology arm corresponds to a nucleic acid sequence homologous to the chromosomal sequence downstream of the integration site. For example, when replacing an exon that harbors one or more mutations, the upstream and downstream homology arms may correspond to the flanking intronic sequences. In embodiments, the homology arms in the exogenous polynucleotide template have 75%, 80%, 85%, 90%, 95%, or 100% sequence identity with the upstream and downstream sequences of the targeted genome sequence. Preferably, the sequence identity is about 95%, 96%, 97%, 98%, 99%, or 100%.

[0996] In certain embodiments, the homology arms of the DNA donor template range from approximately 20 bp to about 5000 bp in length, for example, about 50, 100, 200, 300, 400, 500, 600, 700, 800, 900, 1000, 1100, 1200, 1300, 1400, 1500, 1600, 1700, 1800, 1900, 2000, 2100, 2200, 2300, 2400, 2500, 2600, 2700, 2800, 2900, 3000, 3100, 3200, 3300, 3400, 3500, 3600, 3700, 3800, 3900, 4000, 4100, 4200, 4300, 4400, 4500, 4600, 4700, 4800, 4900, or 5000 bp.

[0997] In some methods, the exogenous DNA template sequence may further comprise a marker. Such a marker may make it easy to screen for targeted integrations. Examples of suitable markers include restriction sites, fluorescent proteins, or selectable markers, such as (antibiotic) resistance genes. The exogenous DNA donor template of the invention can be constructed using recombinant techniques (see, for example, Sambrook et al., 2001 and Ausubel et al., 1996).

[0998] The method of the invention comprises introducing into the cell an exogenous DNA donor template comprising a DNA substitute sequence. The exogenous DNA donor template is an exogenous DNA molecule, not a component of the cell's "original" or "natural" DNA composition, particularly its genome. The DNA donor template comprises or consists of the DNA template sequence.

[0999] In the context of the present invention, the "DNA template sequence," which may also be called "DNA substitute sequence" or "replacement sequence," is a DNA sequence that is introduced into the dsDNA molecule to be modified, replacing a DNA sequence of the dsDNA to be modified that is positioned in proximity to, such as adjacent to and / or between the two single-strand breaks (SSBs).

[1000] In preferred embodiments, the exogenous DNA donor template consists of the DNA substitute sequence. This is possible, for example, if the exogenous DNA donor template is a linear dsDNA molecule, such as a PCR amplification product or a DNA mini-circle. A mini-circle is a plasmid-like circular DNA with all other parts except the sequence of interest removed. Thus, a single cut can linearize a fragment ready to integrate.

[1001] Sequence variants of the claimed nucleic acids, proteins, antibodies, antibody fragments, and / or CARs, for example, those defined by % sequence identity, that maintain similar properties of the invention are also included in the scope of the invention. Such variants, which show alternative sequences, maintain essentially the same properties, such as target specificity, as the specific sequences provided, are known as functional analogs or functionally analogous. Sequence identity relates to the percentage of identical nucleotides or amino acids when carrying out a sequence alignment.

[1002] The recitation “sequence identity,” as used herein, refers to the extent that sequences are identical on a nucleotide-by-nucleotide basis or an amino acid-by-amino acid basis over a window of comparison. Thus, a “percentage of sequence identity” may be calculated by comparing two optimally aligned sequences over the window of comparison, determining the number of positions at which the identical nucleic acid base (e.g., A, T, C, G, I) or the identical amino acid residue (e.g., Ala, Pro, Ser, Thr, Gly, Vai, Leu, He, Phe, Tyr, Trp, Lys, Arg, His, Asp, Glu, Asn, Gin, Cys, and Met) occurs in both sequences to yield the number of matched positions, dividing the number of matched positions by the total number of positions in the window of comparison (i.e. , the window size), and multiplying the result by 100 to yield the percentage of sequence identity. Included are nucleotides and polypeptides having at least about 50%, 55%, 60%, 65%, 70%, 75%, 80%, 85%, 90%, 95%, 96%, 97%, 98%, 99% or 100% sequence identity to any of the reference sequences described herein, typically where the polypeptide variant maintains at least one biological activity of the reference polypeptide.

[1003] It will be appreciated by those of ordinary skill in the art that, as a result of the degeneracy of the genetic code, there are many nucleotide sequences that encode a polypeptide as described herein. Some of these polynucleotides bear minimal homology or sequence identity to the nucleotide sequence of any native gene. Nonetheless, polynucleotides that vary due to differences in codon usage are specifically contemplated by the present invention. Deletions, substitutions, and other changes in sequence that fall under the described sequence identity are also encompassed in the invention.

[1004] Application

[1005] A “chimeric antigen receptor (CAR)” polypeptide comprises an extracellular antigen-binding domain, comprising an antibody or antibody fragment that binds a target antigen, a transmembrane domain, and an intracellular domain. CARs are typically described as comprising an extracellular ectodomain (antigen-binding domain) derived from an antibody and an endodomain comprising signaling modules derived from T-cell signaling proteins.

[1006] In a preferred embodiment, the ectodomain preferably comprises variable regions from the heavy and light chains of an immunoglobulin configured as a single-chain variable fragment (scFv). The scFv is preferably attached to a hinge region that provides flexibility and transduces signals through an anchoring transmembrane moiety to an intracellular signaling domain. The transmembrane domains originate preferably from either CD8a or CD28. In the first generation of CARs, the signaling domain consists of the zeta chain of the TCR complex. The term “generation” refers to the structure of the intracellular signaling domains. Second-generation CARs are equipped with a single costimulatory domain originating from CD28 or 4-1 BB. Third-generation CARs already include two costimulatory domains, e.g., CD28, 4-1 BB, ICOS or 0X40, CDS zeta. The present invention preferably relates to a second or third-generation CAR, although the antigen-binding fragments described herein may be employed in any given CAR format.

[1007] As used herein, the term "chimeric" describes being composed of parts of different proteins or DNAs from different origins.

[1008] The "extracellular antigen-binding domain" or "extracellular binding domain" are used interchangeably and provide a CAR with the ability to bind specifically to the target antigen of interest. The binding domain may be derived either from a natural, synthetic, semi-synthetic, or recombinant source. Preferred are scFv domains.

[1009] "Specific binding" is to be understood by a person skilled in the art, whereby the skilled person is aware of various experimental procedures that can be used to test binding and binding specificity. Methods for determining equilibrium association or equilibrium dissociation constants are known in the art. Cross-reaction or background binding may be inevitable in many protein-protein interactions; this is not to detract from the "specificity" of the binding between CAR and epitope. "Specific binding" describes the binding of an anti-herpes virus antigen-antibody or antigen-binding fragment thereof (or a CAR comprising the same) to said herpes virus antigen at greater binding affinity than background binding. The term "directed against" is also applicable when considering the term "specificity" in understanding the interaction between antibodies and epitopes.

[1010] An "antigen (Ag)" refers to a compound, composition, or substance that can stimulate the production of antibodies or a T cell response in an animal. An "epitope" refers to the region of an antigen to which a binding agent binds. Epitopes can be formed both from contiguous amino acids or noncontiguous amino acids juxtaposed by tertiary folding of a protein. Any suitable antigen can be selected depending on the disease to be treated.

[1011] As used herein, the term "genetically engineered" or "genetically modified" refers to the addition of extra genetic material in the form of DNA or RNA into the total genetic material in a cell. The terms, "genetically modified cells," "modified cells," and, "redirected cells," are used interchangeably.

[1012] An "immune cell" or "immune effector cell" is any cell of the immune system that has one or more effector functions (e.g., cytotoxic cell killing activity, secretion of cytokines, induction of ADCC and / or CDC).

[1013] In embodiments, the immune effector cell is a cytotoxic immune effector cell, preferably a lymphocyte, more preferably a T cell or natural killer (NK) cell.

[1014] Immune effector cells of the invention can be autologous / autogeneic ("self') or non-autologous ("non-self," e.g., allogeneic, syngeneic, or xenogeneic). "Autologous," as used herein, refers to cells from the same subject. "Allogeneic," as used herein, refers to cells of the same species that differ genetically from the cell in comparison. "Syngeneic," as used herein, refers to cells of a different subject that are genetically identical to the cell in comparison. "Xenogeneic," as used herein, refers to cells of a different species to the cell in comparison. In preferred embodiments, the cells of the invention are autologous or allogeneic.

[1015] Illustrative immune effector cells used with the CARs contemplated herein include T lymphocytes. The terms "T cell" or"T lymphocyte" are art-recognized and are intended to include thymocytes, immature T lymphocytes, mature T lymphocytes, resting T lymphocytes, cytokine-induced killer cells (CIK cells), or activated T lymphocytes. Cytokine-induced killer (CIK) cells are typically CDS- and CD56-positive, non-major histocompatibility complex (MHC)-restricted, natural killer (NK)-like T lymphocytes. A T cell can be a T helper (Th; CD4+ T cell) cell, for example, a T helper 1 (Th 1 ) or a T helper s (Th2) cell. The T cell can be a cytotoxic T cell (CTL; CD8+ T cell), CD4+CD8+ T cell, CD4 CDS T cell, or any other subset of T cells. The T cell can be a CD4+ regulatory T cell (T reg; CD4+CD25+Foxp3+ T cell) or a CD8+ regulatory T cells (CD8+CD25+Foxp3+), Other illustrative populations of T cells suitable for use in particular embodiments include naive T cells and memory T cells.

[1016] As would be understood by the skilled person, other cells may also be used as immune effector cells with the CARs described herein. In particular, immune effector cells include NK cells, NKT cells, neutrophils, and macrophages. Immune effector cells also include progenitors of effector cells, wherein such progenitor cells can be induced to differentiate into immune effector cells in vivo or in vitro.

[1017] “Natural Killer (NK) cells,” also known as large granular lymphocytes (LGL), are a vital component of the innate immune system, comprising 5-20% of circulating lymphocytes in humans. Functionally analogous to cytotoxic T cells in adaptive immunity, NK cells respond rapidly to virus-infected cells and tumors, acting approximately three days after infection. Unlike most immune cells, NK cells do not require antibodies or major histocompatibility complex (MHC) for target recognition, allowing for a swift immune response. Identified by the presence of CD56 and absence of CDS, NK cells differentiate from CD127+common innate lymphoid progenitors and mature in various tissues before entering circulation. CD56brightNK cells, predominant in bone marrow and lymphoid tissues, exhibit immunoregulatory roles, while CD56dim NK cells in peripheral blood excel in cell killing. There are various sources from which NK cells can be derived, e.g., peripheral blood mononuclear cells, cord blood, immortalized cell lines, hematopoietic stem and progenitor cells (HSPCs), and induced pluripotent stem cells (iPSCs). All sources can provide clinically meaningful cell doses, are amenable to CAR receptor engineering, and have transitioned into in-human studies. NK cells express distinct receptors, including inhibitory receptors recognizing self-MHC I and activating receptors (Raftery MJ, Ann Rev of Cancer Biol, 2023). Activating molecules may include the natural cytotoxicity receptors, NKG2D, some CD94 / NKG2 complexes, and CD16 (Fcylll). Activating molecules recognize partner ligands on cells except for CD16, a low-affinity Fc receptor responsible for antibody-dependent cellular cytotoxicity (ADCC). Inhibitory molecules may include the killer cell immunoglobulin-like receptor (KIR) family of molecules, some CD94 / NKG2 complexes, and leukocyte Ig-like receptors (LIRs). These receptors govern NK cell activity, adjusting to environmental cues. NK cells contribute to both innate and adaptive immunity, demonstrating antigen-specific immunological memory.

[1018] General remarks

[1019] All words and terms used herein shall have the same meaning commonly given to them by the person skilled in the art unless the context indicates a different meaning. All terms used in the singular shall include the plural of that term and vice versa.

[1020] It will be understood that particular embodiments described herein are shown by way of illustration and not as limitations of the invention. The principal features of this invention can be employed in various embodiments without departing from the scope of the invention. Those skilled in the art will recognize or be able to ascertain, using most routine study, numerous equivalents to the specific procedures described herein. Such equivalents are considered to be within the scope of this invention and are covered by the claims. All publications and patent applications mentioned in the specification indicate the skill level of those skilled in the art to which this invention pertains. All publications and patent applications are herein incorporated by reference to the same extent as if each individual publication or patent application was specifically and individually indicated to be incorporated by reference.

[1021] The use of the word "a" or "an" when used in conjunction with the term "comprising" in the claims and / or the specification may mean "one," but it is also consistent with the meaning of "one or more," "at least one," and "one or more than one." The term "or" in the claims is used to mean "and / or" unless explicitly indicated to refer to alternatives only or the alternatives are mutually exclusive. However, the disclosure supports a definition of only alternatives and "and / or." Throughout this application, where relevant, the term "about" indicates that a value includes the inherent variation of error for the device, the method employed to determine the value or the variation among the study subjects.

[1022] FIGURES

[1023] The invention is demonstrated by way of example in the following figures. The figures are to provide a further description of potentially preferred embodiments that enhance the support of one or more non-limiting embodiments of the invention.

[1024] Brief description of the figures:

[1025] Figure 1 : Enhancement of BEKI using cas target sequences (CTS).

[1026] Figure 2: BEKI mediated simultaneous double knock-in of a truncated CAR into CD3£ and HLA-E into B2M.

[1027] Figure 3: KO-efficiency when performing double nick-mediated TRAC editing using nCas9, ABE or CBE.

[1028] Figure 4: Correlation of Kl-efficiency when using different editors.

[1029] Figure 5: HDR enhancers improve BEKI editing efficiency and enable potent in vivo activity of CD19 CAR T cells.

[1030] Figure 6: BEKI minimizes chromosomal translocations during simultaneous knock-in and multi-gene knockout.

[1031] Figure 7: KO-efficiency when performing double nick-mediated TRAC editing using nCas9, ABE or CBE.

[1032] Figure 8: Correlation of Kl-efficiency when using different editors.

[1033] Figure 9: Impact of potential base edits on BEKI-efficiency using ABE or CBE.

[1034] Figure 10: Gating strategies for flow cytometric analysis of the Kl-efficiency.

[1035] Figure 11 : KO-efficiency of BEKI at the CD3^„ B2M or CD3s locus using ABE.

[1036] Figure 12: BEKI gene editing-efficiency at multiple therapeutically relevant loci can be enhanced using small molecule inhibitors.

[1037] Figure 13: Base editing-mediated splice site disruption of the FKBP12 and RASA2 gene.

[1038] Figure 14: Efficient BEKI engineering of quintuple-edited CAR T cells.

[1039] Figure 15: Kl-efficiency with and without additional retargeting sgRNAs. Figure 16: Implementation of targeted HLA reduction with CBE base editors with BEKI for alternative hypoimmunogenic allogeneic T cell platform.

[1040] Detailed description of the figures:

[1041] Figure 1 : Enhancement of BEKI using cas target sequences (CTS). (A) Cas target sequence (CTS) design for binding of the base editor sgRNA complex to the HDRT in order to enhance transport of the template to the nucleus via the nuclear localization sequence (NLS) (B) BEKI rate using dsDNA, dsCTS Omm, dsC CTS Omm and ssCTS 4mm HDRTs without pharmaceutical enhancement of healthy donors). (C) BEKI rate using dsDNA, dsCTS Omm, dsCTS 4mm, ssCTS Omm 4mm HDRTs with treatment of DNA-PK inhibitor AZD7648 and Pol-Q inhibitor AR healthy donors). Thick lines indicate mean values, error bars indicate Standard error of the mean (SEM).

[1042] Figure 2: BEKI mediated simultaneous double knock-in of a truncated CAR into CD3£ and HLA- E into B2M. (A) BEKI efficiency of a truncCAR into CD3( (left) or HLA-E into B2M (right) after performing CDS^and B2M double KI using BEKI, Cas9 or Cas12a without or with addition of DNA-PK inhibitor AZD7648 and Pol-Q inhibitor ART558. (B) BEKI efficiency of cells carrying both a truncCAR into CDS^and HLA-E into B2M after performing CDS^and B2M double KI using BEKI, Cas9 orCas12a without or with addition of DNA-PK inhibitor AZD7648 and Pol-Q inhibitor ART558. (n=2 healthy donors, 1 pg DNA per electroporation).

[1043] Figure 3: KO-efficiency when performing double nick-mediated TRAC editing using nCas9, ABE or CBE. (A) Schematic overview of the double nick strategy for the insertion of a second generation CD19 CAR into the TRAC locus and the location and direction of the 12 sgRNAs used. (B) Nick distance of the 36 sgRNA combinations targeting opposed DNA strands. Nick distances of PAM-in oriented sgRNA pairs are displayed as negative values. (C) Kl-efficiency using a D10A nCas9 was evaluated on day 4 after electroporation via flow cytometry (n=2 healthy donors). The results are displayed as a heatmap corresponding to B) and as a graph plotting the KI rate against the nick distance. (D) Kl-efficiency using ABE8.20-m was evaluated on day 4 after electroporation via flow cytometry (n=2 healthy donors). The results are displayed as a heatmap corresponding to B) and as a graph plotting the KI rate against the nick distance. (E) Kl-efficiency using a CBE (BE4) were evaluated on day 4 after electroporation via flow cytometry) (n=2 healthy donors). The results are displayed as a heatmap corresponding to B) and as a graph plotting the KI rate against the nick distance.

[1044] Figure 4: Correlation of Kl-efficiency when using different editors. (A) Schematic overview of the double nick strategy for the insertion of a truncated (trunc)CD19 CAR into the CD3£ locus and the location and direction of the 8 sgRNAs evaluated. Heatmaps are showing the nick distance of the 16 sgRNA combinations targeting opposed DNA strands and the Kl-efficiency using ABE8.20-m, evaluated on day 4 after electroporation via flow cytometry (n=4 healthy donors). (B) The same was displayed for the KI of HLA-E into the B2M locus (8 sgRNAs, n=2) and (C) a CD19-specific TRuC into CD3s (5 sgRNAs, n=3 healthy donors). (D) The Kl-rate was normalized to the most efficient KI for each of the 4 tested loci (color) and plotted against the nick distance. The window for optimal editing was manually added. (E) The normalized Kl-rate of editing at the 4 tested loci (color) was plotted against the number of adenines (A) in the editing window (5 and 6 As (n=1 ) are not shown) (F) The normalized Kl-rate (circle size) of 4 tested loci (color) was plotted against the number of adenines (A) in the editing window and the nick distance.

[1045] Figure 5: HDR enhancers improve BEKI editing efficiency and enable potent in vivo activity of CD19 CAR T cells. (A) ABE-mediated Kl-efficiency of a truncCD19 CAR into CD3 ' without HDR enhancers. Z1 was excluded for this experiment due to low initial Kl-rates (n=2 healthy donors). ABE-mediated Kl-efficiency of a truncCD19 CAR into CD3 ' with DNA PK and Pole inhibitors AZD7648 and Novobiocin (n=2 healthy donors). Representative flow cytometry histograms show different editing outcomes when adding DNA PK and Pole inhibitors. Kl-rate of the most efficient combination for CD3^ KI using ABE: Z4+Z7 (n=2 healthy donors). (B) CBE-mediated Kl-efficiency of a truncCD19 CAR into CD3" without HDR enhancers (n=2-6 healthy donors). CBE-mediated Kl- efficiency of a truncCDI 9 CAR into CD3 ' with DNA PK and Pole inhibitors AZD7648 and ART558 (n=2-6 healthy donors). Representative flow cytometry histograms show different editing outcomes when adding DNA PK and Pole inhibitors. Kl-rate of the most efficient combination for CD3^ KI using CBE: Z4+Z8 (n=4 healthy donors). (C) Schematic overview of an acute lymphoblastic leukemia xenograft mouse model using luciferase-labeled Nalm-6 (CD19+) tumor cells. 7 days post Nalm-6 administration, 0.5x106cryopreserved CAR+ T cells were injected systemically. A blood sample was taken from all mice on day 28 after injection of the Nalm-6 tumor cells. Flow cytometric analysis was performed to quantify the frequency of GFP+ Nalm6 cells.

[1046] Figure 6: BEKI minimizes chromosomal translocations during simultaneous knock-in and multi-gene knockout. (A) Schematic overview of the different gene editing strategies targeting the TRAC, RASA2 and FKBP12 gene. (B) Gene editing efficiency of CAR KI and TRAC KO (CDS staining) was determined by flow cytometry. KO- or BE-efficiency of RASA2 and FKBP12 gene were assessed using sanger sequencing followed by indel analysis using ICE CRISPR Analysis. 2025. vS.O. (EditCo Bio) or base editing frequency via EditR. (C) Frequencies of balanced translocations between the on-target genes were determined by ddPCR. They are shown for all six individual translocations and as the sum of all translocations detected in mock, TRF KO (Cas9) an CAR KI + TRF KO (Cas9, Cas12a+ABE, ABE) conditions (n = 3 healthy donors). Statistical analysis of ddPCR was performed using a one-way ANOVA of matched data with Geisser-Greenhouse correction. Multiple comparisons were performed by comparing the mean of each column with the mean of every other column and corrected by the Turkey test. Asterisks represent the following p-values: (not significant [ns], P > .05; *P < .05; **P < .01 ■ *** P < .001).

[1047] Figure 7: KO-efficiency when performing double nick-mediated TRAC editing using nCas9, ABE or CBE. (A) KO-efficiency using a D10A nCas9 was evaluated on day 4 after electroporation via flow cytometry (n=2 healthy donors). The results are displayed as a heatmap corresponding to Fig. 1 B) and as a graphs plotting the KO-rate against the Nick distance or Kl-rate. (B) KO-efficiency using a ABE8.20-m was evaluated on day 4 after electroporation via flow cytometry (n=2 healthy donors). The results are displayed as a heatmap corresponding to Fig. 1 B) and as a graphs plotting the KO-rate against the Nick distance or Kl-rate. (C) KO-efficiency using a CBE (BE4) was evaluated on day 4 after electroporation via flow cytometry (n=2 healthy donors). The results are displayed as a heatmap corresponding to Fig. 1 B) and as a graphs plotting the KO-rate against the Nick distance or Kl-rate. Figure 8: Correlation of Kl-efficiency when using different editors. The mean Kl-efficiency using a D10A nCas9 was plotted against the mean Kl-efficiency using ABE (n = 2 healthy donors). The mean Kl-efficiency using a D10A nCas9 was plotted against the mean Kl-efficiency using CBE (n=2 healthy donors). The mean Kl-efficiency using an ABE was plotted against the mean Kl-efficiency using CBE (n=2 healthy donors).

[1048] Figure 9: Impact of potential base edits on BEKI-efficiency using ABE or CBE. (A) Number of adenines (A) in the editing windows of each sgRNA pair for TRAC editing used in Fig.1 . (B) The Kl- rate (color intensity) using nCas9, ABE or CBE was plotted against the number of A in the editing window and the Nick distance (mean of: n=2 healthy donors). (C) Fold change of the Kl-rate of ABE compared to nCas9 editing using sgRNA pairs in a PAM-out orientation. (D) Number of cytosines (C) in the editing windows of each sgRNA pair for TRAC editing used in Fig.1 . (E) The Kl-rate (colour intensity) using nCas9, ABE or CBE was plotted against the number of C in the editing window and the Nick distance (mean of: n=2 healthy donors). (F) Fold change of the Kl-rate of CBE compared to nCas9 editing using sgRNA pairs in a PAM-out orientation.

[1049] Figure 10: Gating strategies for flow cytometric analysis of the Kl-efficiency. Exemplary flow plots are shown for the KI of a CAR into TRAC, a truncCAR into CD3£, HLA-E into B2M or a TRuC into CD3s (all generated using BEKI using DNA-PKi + PolOi).

[1050] Figure 11 : KO-efficiency of BEKI at the CD3^„ B2M or CD3s locus using ABE. (A) KO- efficiency using a ABE8.20-m was evaluated on day 4 after electroporation via flow cytometry (n = 2- 4 healthy donors). The results are displayed as a heatmap corresponding to Fig. 2. (B) The Kl-rate of BEKI-edited cells at the TRAC, CD3 ',, B2M or CD3s was plotted against the respective KO-rates. The best sgRNA combinations for the 4 loci are highlighted.

[1051] Figure 12: BEKI gene editing-efficiency at multiple therapeutically relevant loci can be enhanced using small molecule inhibitors. (A) BEKI of a truncCD19 CAR into the CD3T locus w / o orwith different DNA-PK inhibitors (Alt-R HDR enhancer V2, AZD7648, Nedisertib (M3814)) for the sgRNA combinations shown in Fig.2 (n=2 healthy donors). (B) Kl-rate of the most efficient combination for ABE-mediated KI into the TRAC, CD3 ', B2M or CD3s locus (n=2-4 healthy donors). Representative flow cytometry histograms show editing outcomes when adding DNA PK and Pol9 inhibitors.

[1052] Figure 13: Base editing-mediated splice site disruption of the FKBP12 and RASA2 gene. (A) BE-efficiency of FKBP12 was assessed using sanger sequencing followed by quantification of the base editing frequency via EditR. The mock electroporated or base edited cells were stimulated using aCD3 / aCD28 antibodies with orw / o Tacrolimus (Tac, final concentration of 6 ng / mL) before flow cytometric analysis of the TNFa and IFNr expression of CD4 and CDS T cells. The frequency of cells expressing cytokines in the presence of Tac was normalized to the frequency w / o Tac treatment. (B) BE-efficiency of RASA2 was assessed using sanger sequencing followed by quantification of the base editing frequency via EditR. The mock electroporated or base edited cells were stimulated using aCD3 / aCD28 antibodies with orw / o Tacrolimus (Tac, final concentration of 6 ng / mL) before flow cytometric analysis of the TNFa and IFNr expression of CD4 and CDS T cells . Unstimulated cells are included as a control (lighter color). (C) Gene edited CD19 CAR T cells or mock electroporated T cells were stimulated using CD19+ Naim 6 tumor cells in the presence or absence of Tac. Flow cytometric analysis of the TNFa and IFNr expression of CD4 and CDS T cells was performed after 6 hours.

[1053] Figure 14: Efficient BEKI engineering of quintuple-edited CAR T cells. (A) Schematic overview of the different gene editing strategies targeting the TRAC, RASA2,FKBP12, CIITA and B2M genes. (B) Gene editing efficiency of CAR KI and TRAC (CDS), B2M (HLA-A, B,C) and CIITA (HLA-DR,DP,DQ) KO was determined by flow cytometry (n=3 healthy donors). (C) Allo-specific CD8+ T cells were generated against another donor by adding irradiated CDS- cells twice. The allo-specific cells were co-cultured with CFSE-labled mock electroporated or gene edited cells from the donor used for stimulation at effectorto target ratios of 0.5:1 , 1 :1 and 2:1 (n=1 healthy donor). The number of CFSE+ cells was determined and the elimination frequency quantified using wells containing only target cells.

[1054] Figure 15: Kl-efficiency with and without additional retargeting sgRNAs. Kl-efficiency of the BEKI- system with (B2 + B2retar.) and without additional retargeting gRNAs (B2exa) was evaluated by flow cytometry.

[1055] Figure 16: Implementation of targeted HLA reduction with CBE base editors with BEKI for alternative hypoimmunogenic allogeneic T cell platform. (A) HLA class-l KOs were performed using CBE4 to introduce premature stop codons. sgRNAs were selected to target HLA-A, B and / or C in T cell of healthy donors with specific HLA-alleles. (B) In an allo-specific T cell cytotoxicity assay it was evaluated if elimination of HLA-A, HLA-C or HLA-A, B,C using gene editing protected the cells from allo-specific T cell cytotoxixity compared to mock electroporated T cells after 18 hours of co-culture with alloreactive CD8+ cells in different effector to target cell ratios. (C) In an NK cell cytotoxicity assay the elimination of HLA-edited cells after 18 hours of co-culture with primary NK cells in different effector to target cell ratios was assessed.

[1056] EXAMPLES

[1057] The invention is demonstrated by way of the examples disclosed below. The examples provide technical support for and a more detailed description of potentially preferred, non-limiting embodiments of the invention.

[1058] Summary of the Examples

[1059] In order to demonstrate the functionality and beneficial properties of the in vitro method for modifying double-stranded DNA (dsDNA) in a eukaryotic cell described herein, the following examples are to be considered:

[1060] T cell isolation and cell culture

[1061] In vitro transctiption for mRNA production

[1062] Design of DNA templates for homology-directed insertion of transgenes

[1063] Generation of HDRT

[1064] HDR-templates with Cas target sequences

[1065] Gene editing Flow cytometry

[1066] Intracellular cytokine detection

[1067] Digital droplet polymerase chain reaction (ddPCR) for translocation quantification

[1068] Sanger sequencing - indel and base editing quantification

[1069] Data analysis, statistics, and presentation

[1070] Combination of two sgRNAs and a nickase Cas9 base editor enables knock-in of transgenes

[1071] Base editor knock-ins require distinct sgRNA designs for efficient editing with both adenine base editors and cytosine base editors

[1072] BEKI can be designed for knock-in of different transgenes and target sites and can be optimized to provide simultaneous knock-out of the endogenous gene at the integration locus

[1073] BEKI rates are increased by exposing cells to DNA repair modulating agents that inhibit non- homologous end joining and microhomology-mediated end joining

[1074] BEKI rates can be enhanced by modification of the non-viral templates with Cas-target sequence motifs

[1075] BEKI can be combined with additional base edits to force gene knock-outs or to induce gain- of-function or loss-of-function mutations alongside transgene integration

[1076] Combining BEKI with additional simultaneous edits in a single transfection prevents translocations

[1077] BEKI can be used to integrate two transgenes at different loci simultaneously while reducing translocations between the different integration sites

[1078] Example 1: T cell isolation and cell culture

[1079] Peripheral blood mononuclear cells (PBMCs) were isolated from the peripheral blood of healthy donors via density gradient centrifugation using Pancoll separation solution (Pan Biotech) in 50mL Leucosep Tubes (Greiner).

[1080] Prior to T cell seeding, 24-well tissue culture plates were coated with 1 pg / mL each of anti-CD3 (Invitrogen) and anti-CD28 (BioLegend) in 0.5 mL / well sterile water (Ampuwa), sealed with parafilm, and incubated at 4 °C for 24 h. CD3+ T cells were positively enriched using magnetic column enrichment with human CDS microbeads according to the manufacturer’s recommendations (LS columns, Miltenyi Biotec). In brief, cells were resuspended in cold MACS buffer (PBS + 2% heat- inactivated FCS + 2 mM EDTA), incubated with anti-CD3-conjugated magnetic beads for 15 min at 4 °C, washed, filtered (30 pm), and separated using LS columns (Miltenyi Biotec). Isolated cells were counted and seeded at 0.75 x 1 O6cells / mL into pre-washed aCD3 / aCD28-coated wells.

[1081] T cells were cultured at 37 °C and 5% CO2in T cell medium containing a 1 :1 mixture of advanced RPMI (Gibco, Thermo Fisher Scientific) and Click’s (Fujifilm Irvine Scientific, Santa Ana, CA) media supplemented with 10% inactivated fetal calf serum (FCS) (Biochrom), 1 % Glutamax (Thermo Fisher Sci...

Claims

AMENDED CLAIMS received by the International Bureau on 31 March 2026 (31.03.2026)1. An in vitro method for modifying double stranded DNA (dsDNA) in a eukaryotic cell, the method comprising: a) introducing a base editor system into the cell, said system comprising i. a base editor, comprising an RNA guided DNA nickase linked to a singlestranded DNA nucleobase modifying enzyme, or a nucleic acid encoding said base editor, and ii. at least two guide RNA (sgRNA) molecules, comprising a first sgRNA that hybridizes to a first target sequence in the dsDNA, and a second sgRNA that hybridizes to a second target sequence in an opposing strand of the dsDNA, b) generating at least two single-strand breaks (SSBs), in opposing strands, of the dsDNA, iii. wherein a first base editor complex, comprising the first sgRNA, introduces a nick cleavage of one strand of the dsDNA in the first target sequence, and a second base editor complex, comprising the second sgRNA, introduces a nick cleavage of the opposite strand of the dsDNA in the second target sequence, thereby producing a 5’ or 3’ overhang, c) inserting an exogenous DNA repair template sequence in proximity to the two single strand breaks (SSBs), wherein the base editor system and the exogenous DNA repair template sequence are introduced into the cell at the same time using a non-viral method.

2. The method according to claim 1 , wherein the exogenous DNA repair template sequence encodes a recombinant antigen-specific targeting construct, such as a chimeric antigen receptor (CAR) or T cell receptor (TCR), or a construct configured to enhance the persistence and / or potency of cells for therapeutic effects.

3. The method according to any one of the preceding claims, wherein the eukaryotic cell is an immune cell and / or an effector cell, such as a Treg, or a cytotoxic immune effector cell, preferably a T cell, or NK cell.

4. The method according to the preceding claim, wherein the eukaryotic cell is a T or NK cell and the first and second target sequences of the dsDNA are in an endogenous genomic dsDNA encoding a T cell receptor component, such as in the TRAC locus, CD3C locus or CD3s locus, or a major histocompatibility complex component, such as in the B2M locus.

5. The method according to any one of the preceding claims, wherein the exogenous DNA repair template sequence comprises at least one sequence (homology arm) of at least 45 bp, preferably 400 bp, with at least 90%, preferably 95%, more preferably 100%,sequence identity to a region of the dsDNA sequence, wherein said homology arm is positioned to hybridize to the dsDNA in proximity to the two single strand breaks (SSBs).

6. The method according to any one of the preceding claims, comprising in step a) introducing one or more additional sgRNA sequences, comprising a third sgRNA that hybridizes to a third target sequence in the dsDNA, and forming a third base editor complex, that introduces a nick cleavage of one strand of the dsDNA and a nucleobase modification in the third target sequence.

7. The method according to the preceding claim, wherein: the third target sequence of the dsDNA is a component of a gene, such as a regulatory sequence, splice donor, splice acceptor or other sequence associated with expression of said gene, and wherein said modification of the third target sequence leads to disruption of expression of said gene, and / or wherein the third target sequence of the dsDNA is located in the open reading frame (ORF) of a gene, associated with a functional domain of said gene, and wherein said modification of the third target sequence leads to an amino acid change leading to a functional alteration of said gene.

8. The method according to the preceding claim, wherein the expression of multiple genes is disrupted (multiple gene knockouts) with multiple additional sgRNA sequences, simultaneously with the insertion of an exogenous DNA repair template sequence.

9. The method according to any one of the preceding claims, wherein the single-stranded DNA nucleobase modifying enzyme is an adenosine deaminase, a cytidine deaminase, cytosine-DNA glycolase, thymidine-DNA glycolase or fusion constructs containing deaminases and glycolases, such as a cytidine deaminase with a uracil glycosylase.

10. The method according to any one of the preceding claims, wherein the base editor is a Cytosine Base Editor (CBE) and / or Adenine Base Editor (ABE) and / or a Glycosylase Base Editor (GBE).11 . The method according to any one of the preceding claims, wherein: the base editor is an ABE, and the first and / or second sgRNA lacks adenine bases in the editing window and / or wherein the base editor is a CBE, and the first and / or second sgRNA lacks cytosine bases in the editing window, or the base editor is an ABE, and the first and second sgRNA together have a total of less than 6 adenine bases in the editing window and / or wherein the base editor is a CBE, and the first and second sgRNA together have a total of less than 6 cytosine bases in the editing window.

12. The method according to any one of the preceding claims, comprising introducing to said cell one or more DNA repair modulators and / or contacting said cell with one or more DNA-dependent protein kinase PK (DNA-PK) inhibitors or DNA PolQ inhibitors.

13. The method according to any one of the preceding claims, wherein:the distance between the two single strand breaks on the dsDNA is more than 29 bases, and / or the distance between the two single strand breaks on the dsDNA is less than 250 bases, preferably less than 200 bases.

14. A genetically modified eukaryotic cell obtained by the method according to any one of the preceding claims.

15. The genetically modified eukaryotic cell according to the preceding claim, comprising at least one insertion of an exogenous DNA donor template sequence and at least one base edit in the dsDNA produced by the base editor system in proximity to said exogenous DNA donor template sequence.

16. A kit for use in the in vitro method for modifying double stranded DNA (dsDNA) according to any one of claims 1 to 16, the kit comprising: a. a base editor system, said system comprising i. a base editor, comprising an RNA guided DNA nickase linked to a singlestranded DNA nucleobase modifying enzyme, or a nucleic acid encoding said base editor, and ii. at least two guide RNA (sgRNA) sequences, comprising a first sgRNA that hybridizes to a first target sequence in the dsDNA, and a second sgRNA that hybridizes to a second target sequence in the dsDNA, wherein a first base editor complex, comprising the first sgRNA, is configured to introduce a nick cleavage of one strand of the dsDNA in the first target sequence, and a second base editor complex, comprising the second sgRNA, is configured to introduce a nick cleavage of the opposite strand of the dsDNA in the second target sequence, thereby producing a 5’ or 3’ overhang, and b. an exogenous DNA repair template that hybridizes to the dsDNA in proximity to one or both single strand breaks (SSBs), wherein the base editor system and the exogenous DNA repair template sequence are configured to be introduced into the cell at the same time using a non-viral method, c. and optionally one or more DNA repair modulators.

Citation Information

Patent Citations

  • Reprogramming immune cells by targeted integration of ZETA-deficient chimeric antigen receptor transgenes

    WO2022136551A1

  • Enhancing efficiency of targeted gene knockin by base editors

    WO2022159753A1