OMNI XL1-22 CRISPR nuclease
OMNI CRISPR nucleases with modified catalytic sites and guide RNA targeting enhance genome editing precision, addressing sequence and delivery limitations of current CRISPR technologies.
Patent Information
- Application Number
- JP2025539451
- Authority / Receiving Office
- JP · JP
- Patent Type
- Applications
- Current Assignee / Owner
- Priority Date
- 2023-01-03
- Filing Date
- 2024-01-03
- Publication Date
- 2026-01-16
AI Technical Summary
Current CRISPR nucleases face limitations in sequence specificity, expression, and delivery, which restrict their applications in genome engineering due to PAM site restrictions and potential pre-existing immunity, hindering their in vivo applicability.
Development of OMNI CRISPR nucleases with modified catalytic sites for nickase or dead nuclease activity, combined with guide RNA molecules to target specific DNA sequences, and fusion with DNA-modifying enzymes for precise genome editing and mutation correction.
Enhances the specificity and versatility of genome editing by allowing targeted mutations and modifications, overcoming sequence and delivery constraints, and enabling applications in genome engineering and diagnostics.
Smart Images

Figure 2026501682000001 
Figure 2026501682000002 
Figure 2026501682000003
Abstract
Description
[Technical Field]
[0001] This application claims the benefit of U.S. Provisional Application No. 63 / 478,292, filed January 3, 2023, the contents of which are incorporated herein by reference. This application also references various publications, the disclosures of which are incorporated herein by reference to, inter alia, further describe techniques and features that may be used in the present invention.
[0002] Sequence Listing Reference This application was created on December 15, 2022 on an IBM PC machine using an operating system compatible with MS-Windows®, and incorporates by reference the nucleotide sequence in the XML file with file name "230103_91 820-PRO_Sequence_Listing_AWG.xml", 522 KB in size, filed as part of this application on January 3, 2024.
[0003] Technical Field The present invention relates, inter alia, to compositions and methods for genome editing. [Background technology]
[0004] The clustered regularly interspaced short palindromic repeats (CRISPR) system in bacterial and archaeal adaptive immunity exhibits extreme diversity in protein composition and genomic locus structure. CRISPR systems have become important tools for research and genome engineering. Nevertheless, many details of CRISPR systems remain unknown, and the application of CRISPR nucleases may be limited by sequence specificity, expression, or delivery. Different CRISPR nucleases have diverse characteristics, including size, PAM site, on-target activity, specificity, cleavage patterns (e.g., blunt ends, sticky ends), and prominent patterns of indel formation after cleavage. Combinations of diverse properties may be useful for various applications. For example, some CRISPR nucleases can target specific genomic loci, while others cannot due to PAM site restrictions. Furthermore, some currently used CRISPR nucleases exhibit pre-existing immunity, potentially limiting their in vivo applicability. See Charlesworth et al., Nature Medicine (2019) and Wagner et al., Nature Medicine (2019). Therefore, the discovery, application, and improvement of novel CRISPR nucleases is important. Summary of the Invention
[0005] Disclosed herein are compositions and methods that can be used for genome engineering, epigenome engineering, genome targeting, cellular genome editing, and / or in vitro diagnostics.
[0006] The disclosed compositions may be used to modify genomic DNA sequences. As used herein, genomic DNA refers to linear and / or chromosomal DNA and / or plasmid or other extrachromosomal DNA sequences present in a cell of interest. In some embodiments, the cell of interest is a eukaryotic cell. In some embodiments, the cell of interest is a prokaryotic cell. In some embodiments, the method generates a double-strand break (DSB) at a predetermined target site in the genomic DNA sequence, resulting in a mutation, addition, and / or deletion of the DNA sequence at the target site in the genome.
[0007] Thus, in some embodiments, the composition contains a clustered regularly interspaced short palindromic repeat (CRISPR) nuclease. In some embodiments, the CRISPR nuclease is a CRISPR-associated protein.
[0008] OMNI CRISPR nuclease Embodiments of the invention provide CRISPR nucleases referred to as "OMNI" nucleases, which are shown in Table 1.
[0009] The present invention provides a method for modifying a nucleotide sequence at a target site in the genome of a mammalian cell, comprising: (i) a composition containing a nucleic acid molecule comprising a sequence encoding a CRISPR nuclease having at least 95% identity to a sequence selected from the group consisting of the amino acid sequences set forth in SEQ ID NOs: 1 to 22, or a CRISPR nuclease having a sequence at least 95% identity to a sequence selected from the group consisting of the nucleic acid sequences set forth in SEQ ID NOs: 23 to 44; and (ii) a DNA-targeting RNA molecule comprising a nucleotide sequence complementary to the sequence of the target DNA, or a DNA polynucleotide encoding the DNA-targeting RNA molecule, into the cell.
[0010] This invention is One or more RNA molecules comprising a guide sequence portion linked to direct repeats capable of hybridizing to a target sequence, or one or more nucleotide sequences encoding said one or more RNA molecules; and a CRISPR nuclease comprising an amino acid sequence having at least 95% identity to a sequence selected from the group consisting of the amino acid sequences set forth in SEQ ID NOs: 1 to 22, or a nucleic acid molecule comprising a sequence encoding the CRISPR nuclease; Also provided are non-naturally occurring compositions containing a CRISPR-associated system comprising: wherein the one or more RNA molecules hybridize to the target sequence, the target sequence being adjacent to the 3' end of a complementary sequence of a protospacer adjacent motif (PAM), and the one or more RNA molecules form a complex with an RNA-guided nuclease.
[0011] This invention is A CRISPR nuclease comprising a sequence having at least 95% identity to a sequence selected from the group consisting of the amino acid sequences set forth in SEQ ID NOs: 1 to 22, or a nucleic acid molecule comprising a sequence encoding the CRISPR nuclease; and a nucleotide sequence of a nuclease-binding RNA capable of interacting with / binding to the CRISPR nuclease; and a nucleotide sequence of the DNA-targeting RNA that comprises a sequence complementary to a sequence in the target DNA sequence; one or more RNA molecules, or one or more DNA polynucleotides encoding said one or more RNA molecules, comprising at least one of: Also provided are non-naturally occurring compositions containing a CRISPR-associated system comprising: Here, the CRISPR nuclease can complex with the one or more RNA molecules to form a complex that can hybridize to the target DNA sequence. DETAILED DESCRIPTION OF THE INVENTION
[0012] Detailed Description In some aspects of the invention, the disclosed compositions contain a clustered regularly interspaced short palindromic repeat (CRISPR) nuclease and / or a nucleic acid molecule comprising a sequence encoding the same.
[0013] Table 1 shows novel CRISPR nucleases and one or more substitution positions within the nuclease that convert the nuclease into a nickase or a dead nuclease. For example, the catalytic site of a CRISPR nuclease of the invention may be modified so that the nuclease has nickase activity and can cleave single-stranded DNA. Alternatively, the catalytic site of a CRISPR nuclease of the invention may be modified so that the nuclease does not have nuclease activity, i.e., is inactivated. Although reference is made to CRISPR nucleases herein, these nucleases may be modified to have nickase activity (i.e., nucleases that cleave single-stranded DNA rather than double-stranded DNA) or to have no nuclease activity (i.e., a dead nuclease).
[0014] Table 2 lists crRNA, tracrRNA, and single guide RNA (sgRNA) sequences, as well as portions of crRNA, tracrRNA, and sgRNA sequences compatible with each CRISPR nuclease. Thus, a crRNA molecule capable of binding to and targeting an OMNI-nuclease shown in Table 2 as part of a crRNA:tracrRNA complex may comprise a crRNA sequence shown in Table 2. Similarly, a tracrRNA molecule capable of binding to and targeting an OMNI-nuclease shown in Table 2 as part of a crRNA:tracrRNA complex may comprise a tracrRNA sequence shown in Table 2. Additionally, a single guide RNA molecule capable of binding to and targeting an OMNI-nuclease shown in Table 2 may comprise a sequence shown in Table 2.
[0015] For example, a crRNA molecule for OMNI-XL-12 nuclease (SEQ ID NO: 12) may comprise the sequence set forth in any one of SEQ ID NOs: 232-235; a tracrRNA molecule for OMNI-XL-12 nuclease may comprise the sequence set forth in any one of SEQ ID NOs: 236-244 and 246-249; and a sgRNA molecule for OMNI-XL-12 nuclease may comprise the sequence set forth in any one of SEQ ID NOs: 231-249. Other crRNA, tracrRNA, or sgRNA molecules for OMNI nuclease can be derived in the same manner from the sequences set forth in Table 2.
[0016] These nucleases can utilize guide RNA molecules to target desired target DNA sequences. The nuclease-guide complex delivers any molecule bound to the complex to the target site. Thus, this disclosure also contemplates fusion proteins comprising a CRISPR nuclease and a DNA-modifying domain (e.g., a deaminase, nuclease, nickase, recombinase, methyltransferase, methylase, acetylase, acetyltransferase, transcriptional activator, or transcriptional repressor domain), and the use of such fusion proteins to modify disease-associated mutations in a genome (e.g., the genome of a human subject) or to generate mutations in a genome (e.g., the human genome) to reduce or prevent gene expression.
[0017] In some embodiments, the CRISPR nuclease of the present invention may be fused with a protein having enzymatic activity. In some embodiments, the enzymatic activity modifies target DNA. In some embodiments, the enzymatic activity is nuclease activity, methyltransferase activity, demethylase activity, DNA repair activity, DNA damage activity, deamination activity, dismutase activity, alkylation activity, depurination activity, oxidation activity, pyrimidine dimer formation activity, integrase activity, transposase activity, recombinase activity, polymerase activity, ligase activity, helicase activity, photolyase activity, or glycosylase activity. In some cases, the enzymatic activity is nuclease activity. In some cases, the nuclease activity induces double-strand breaks in target DNA. In some cases, the enzymatic activity modifies target polypeptides associated with target DNA. In some cases, the enzymatic activity is a methyltransferase activity, a demethylase activity, an acetyltransferase activity, a deacetylase activity, a kinase activity, a phosphatase activity, a ubiquitin ligase activity, a deubiquitinating activity, an adenylating activity, a deadenylating activity, a sumoylating activity, a desumoylating activity, a ribosylation activity, a deribosylation activity, a myristoylating activity, or a demyristoylating activity. In some cases, the target polypeptide is a histone and the enzymatic activity is a methyltransferase activity, a demethylase activity, an acetyltransferase activity, a deacetylase activity, a kinase activity, a phosphatase activity, a ubiquitin ligase activity, or a deubiquitinating activity.
[0018] Thus, any one of the CRISPR nucleases, nickases, or dead nucleases may be fused (e.g., directly or via a linker) to another DNA-modifying or DNA-modifying enzyme, including, but not limited to, a deaminase, a reverse transcriptase (e.g., for use in prime editing, see Anzalone et al. (2019)), an enzyme that alters the methylation state of DNA (e.g., a methyltransferase), or a modifier of histones (e.g., a histone acetyltransferase). Indeed, the OMNI-50 nucleases, nickases, and dead nucleases described herein may be fused to a DNA-modifying enzyme or its effector domain. Examples of DNA-modifying enzymes include, but are not limited to, deaminases, nucleases, nickases, recombinases, methyltransferases, methylases, acetylases, acetyltransferases, reverse transcriptases, helicases, integrases, ligases, transposases, demethylases, phosphatases, transcriptional activators, or transcriptional repressors. In some embodiments, the CRISPR nucleases of the present invention are fused to proteins with enzymatic activity. In some embodiments, the enzymatic activity modifies a target DNA molecule. The CRISPR nucleases or fusion proteins thereof described herein may be used to correct or generate one or more mutations in a gene associated with a disease, or to increase, correct, reduce, or prevent the expression of a gene.
[0019] The present invention provides a non-naturally occurring composition comprising a CRISPR nuclease comprising a sequence having at least 90% identity to a sequence selected from the group consisting of the amino acid sequences set forth in SEQ ID NOs: 12, 16, 18, 19, 1-11, 13-15, 17, and 20-22, or a nucleic acid molecule comprising a sequence encoding the CRISPR nuclease.
[0020] In some embodiments, the composition further comprises one or more RNA molecules or a DNA polynucleotide encoding any one of said one or more RNA molecules, wherein said one or more RNA molecules and said CRISPR nuclease are not found together in nature, and said one or more RNA molecules are configured to form a complex with the CRISPR nuclease and / or said one or more RNA molecules are configured to target the complex to a target site.
[0021] In some embodiments, the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 1, and at least one RNA molecule comprises a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 49-64.
[0022] In some embodiments, the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 1, and the at least one RNA molecule is a CRISPR RNA (crRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 50-53.
[0023] In some embodiments, the composition further comprises a trans-activating CRISPR RNA (tracrRNA) molecule comprising a sequence set forth in the group consisting of SEQ ID NOs: 54-64.
[0024] In some embodiments, the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 1, and the at least one RNA molecule is a single guide RNA (sgRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 49-64.
[0025] In some embodiments, the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 2, and at least one RNA molecule comprises a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 65-80.
[0026] In some embodiments, the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO:2, and the at least one RNA molecule is a CRISPR RNA (crRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs:66-69.
[0027] In some embodiments, the composition further comprises a trans-activating CRISPR RNA (tracrRNA) molecule comprising a sequence set forth in the group consisting of SEQ ID NOs: 70-80.
[0028] In some embodiments, the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 2, and the at least one RNA molecule is a single guide RNA (sgRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 65-80.
[0029] In some embodiments, the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 3, and at least one RNA molecule comprises a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 81-96.
[0030] In some embodiments, the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO:3, and the at least one RNA molecule is a CRISPR RNA (crRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs:82-85.
[0031] In some embodiments, the composition further comprises a trans-activating CRISPR RNA (tracrRNA) molecule comprising a sequence set forth in the group consisting of SEQ ID NOs: 86-96.
[0032] In some embodiments, the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 3, and the at least one RNA molecule is a single guide RNA (sgRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 81-96.
[0033] In some embodiments, the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 4, and at least one RNA molecule comprises a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 97-112.
[0034] In some embodiments, the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO:4, and the at least one RNA molecule is a CRISPR RNA (crRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs:98-101.
[0035] In some embodiments, the composition further comprises a trans-activating CRISPR RNA (tracrRNA) molecule comprising a sequence set forth in the group consisting of SEQ ID NOs: 102-112.
[0036] In some embodiments, the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO:4, and the at least one RNA molecule is a single guide RNA (sgRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs:97-112.
[0037] In some embodiments, the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 5, and at least one RNA molecule comprises a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 113-128.
[0038] In some embodiments, the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO:5, and the at least one RNA molecule is a CRISPR RNA (crRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs:114-117.
[0039] In some embodiments, the composition further comprises a trans-activating CRISPR RNA (tracrRNA) molecule comprising a sequence set forth in the group consisting of SEQ ID NOs: 118-128.
[0040] In some embodiments, the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO:5, and the at least one RNA molecule is a single guide RNA (sgRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs:113-128.
[0041] In some embodiments, the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 6, and at least one RNA molecule comprises a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 129-144.
[0042] In some embodiments, the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO:6, and the at least one RNA molecule is a CRISPR RNA (crRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs:130-133.
[0043] In some embodiments, the composition further comprises a trans-activating CRISPR RNA (tracrRNA) molecule comprising a sequence set forth in the group consisting of SEQ ID NOs: 134-144.
[0044] In some embodiments, the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 6, and the at least one RNA molecule is a single guide RNA (sgRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 129-144.
[0045] In some embodiments, the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 7, and at least one RNA molecule comprises a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 145-163.
[0046] In some embodiments, the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 7, and the at least one RNA molecule is a CRISPR RNA (crRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 146-149.
[0047] In some embodiments, the composition further comprises a trans-activating CRISPR RNA (tracrRNA) molecule comprising a sequence set forth in the group consisting of SEQ ID NOs: 150-160, 162, and 163.
[0048] In some embodiments, the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 7, and the at least one RNA molecule is a single guide RNA (sgRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 145-163.
[0049] In some embodiments, the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 8, and at least one RNA molecule comprises a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 164-179.
[0050] In some embodiments, the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 8, and the at least one RNA molecule is a CRISPR RNA (crRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 165-168.
[0051] In some embodiments, the composition further comprises a trans-activating CRISPR RNA (tracrRNA) molecule comprising a sequence set forth in the group consisting of SEQ ID NOs: 169-179.
[0052] In some embodiments, the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 8, and the at least one RNA molecule is a single guide RNA (sgRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 164-179.
[0053] In some embodiments, the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 9, and at least one RNA molecule comprises a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 180-195.
[0054] In some embodiments, the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO:9, and the at least one RNA molecule is a CRISPR RNA (crRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs:181-184.
[0055] In some embodiments, the composition further comprises a trans-activating CRISPR RNA (tracrRNA) molecule comprising a sequence set forth in the group consisting of SEQ ID NOs: 185-195.
[0056] In some embodiments, the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 9, and the at least one RNA molecule is a single guide RNA (sgRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 180-195.
[0057] In some embodiments, the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 10, and at least one RNA molecule comprises a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 196-211.
[0058] In some embodiments, the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 10, and the at least one RNA molecule is a CRISPR RNA (crRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 197-200.
[0059] In some embodiments, the composition further comprises a trans-activating CRISPR RNA (tracrRNA) molecule comprising a sequence set forth in the group consisting of SEQ ID NOs: 201-211.
[0060] In some embodiments, the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 10, and the at least one RNA molecule is a single guide RNA (sgRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 196-211.
[0061] In some embodiments, the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 11, and at least one RNA molecule comprises a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 212-230.
[0062] In some embodiments, the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 11, and the at least one RNA molecule is a CRISPR RNA (crRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 213-216.
[0063] In some embodiments, the composition further comprises a trans-activating CRISPR RNA (tracrRNA) molecule comprising a sequence set forth in the group consisting of SEQ ID NOs: 217-225 and 227-230.
[0064] In some embodiments, the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 11, and the at least one RNA molecule is a single guide RNA (sgRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 212-230.
[0065] In some embodiments, the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 12, and at least one RNA molecule comprises a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 231-249.
[0066] In some embodiments, the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 12, and the at least one RNA molecule is a CRISPR RNA (crRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 232-235.
[0067] In some embodiments, the composition further comprises a trans-activating CRISPR RNA (tracrRNA) molecule comprising a sequence set forth in the group consisting of SEQ ID NOs: 236-244 and 246-249.
[0068] In some embodiments, the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 12, and the at least one RNA molecule is a single guide RNA (sgRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 231-249.
[0069] In some embodiments, the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 13, and at least one RNA molecule comprises a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 250-270.
[0070] In some embodiments, the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 13, and the at least one RNA molecule is a CRISPR RNA (crRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 251-254.
[0071] In some embodiments, the composition further comprises a trans-activating CRISPR RNA (tracrRNA) molecule comprising a sequence set forth in the group consisting of SEQ ID NOs: 255-265 and 267-270.
[0072] In some embodiments, the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 13, and at least one RNA molecule is a single guide RNA (sgRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 250-270.
[0073] In some embodiments, the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 14, and at least one RNA molecule comprises a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 271-286.
[0074] In some embodiments, the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 14, and the at least one RNA molecule is a CRISPR RNA (crRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 272-275.
[0075] In some embodiments, the composition further comprises a trans-activating CRISPR RNA (tracrRNA) molecule comprising a sequence set forth in the group consisting of SEQ ID NOs: 276-286.
[0076] In some embodiments, the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 14, and the at least one RNA molecule is a single guide RNA (sgRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 271-286.
[0077] In some embodiments, the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 15, and at least one RNA molecule comprises a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 287-302.
[0078] In some embodiments, the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 15, and the at least one RNA molecule is a CRISPR RNA (crRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 288-291.
[0079] In some embodiments, the composition further comprises a trans-activating CRISPR RNA (tracrRNA) molecule comprising a sequence set forth in the group consisting of SEQ ID NOs: 292-302.
[0080] In some embodiments, the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 15, and the at least one RNA molecule is a single guide RNA (sgRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 287-302.
[0081] In some embodiments, the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 16, and at least one RNA molecule comprises a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 303-318.
[0082] In some embodiments, the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 16, and the at least one RNA molecule is a CRISPR RNA (crRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 304-307.
[0083] In some embodiments, the composition further comprises a trans-activating CRISPR RNA (tracrRNA) molecule comprising a sequence set forth in the group consisting of SEQ ID NOs: 308-318.
[0084] In some embodiments, the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 16, and the at least one RNA molecule is a single guide RNA (sgRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 303-318.
[0085] In some embodiments, the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 17, and at least one RNA molecule comprises a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 319-332.
[0086] In some embodiments, the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 17, and the at least one RNA molecule is a CRISPR RNA (crRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 320-323.
[0087] In some embodiments, the composition further comprises a trans-activating CRISPR RNA (tracrRNA) molecule comprising a sequence set forth in the group consisting of SEQ ID NOs: 324-332.
[0088] In some embodiments, the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 17, and the at least one RNA molecule is a single guide RNA (sgRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 319-332.
[0089] In some embodiments, the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 18, and at least one RNA molecule comprises a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 333-353.
[0090] In some embodiments, the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 18, and the at least one RNA molecule is a CRISPR RNA (crRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 334-337.
[0091] In some embodiments, the composition further comprises a trans-activating CRISPR RNA (tracrRNA) molecule comprising a sequence set forth in the group consisting of SEQ ID NOs: 338-348 and 350-353.
[0092] In some embodiments, the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 18, and the at least one RNA molecule is a single guide RNA (sgRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 333-353.
[0093] In some embodiments, the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 19, and at least one RNA molecule comprises a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 354-369.
[0094] In some embodiments, the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 19, and the at least one RNA molecule is a CRISPR RNA (crRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 355-358.
[0095] In some embodiments, the composition further comprises a trans-activating CRISPR RNA (tracrRNA) molecule comprising a sequence set forth in the group consisting of SEQ ID NOs: 359-369.
[0096] In some embodiments, the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 19, and the at least one RNA molecule is a single guide RNA (sgRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 354-369.
[0097] In some embodiments, the amino acid sequence of the CRISPR nuclease comprises a conserved insertion sequence consisting of 343 consensus sites, each of which may be an amino acid or gap such that removing the gap and joining the amino acids results in the full amino acid sequence of the conserved insertion region; Consensus site 1 of the conserved insertion region is an asparagine (N) residue; Consensus site 4 of the conserved insertion region is a leucine (L) residue; The consensus site 45 of the conserved insertion region is a lysine (K) residue; The consensus site 50 of the conserved insertion region is a proline (P) residue; The consensus site 61 of the conserved insertion region is a valine (V) residue; The consensus site 63 in the conserved insertion region is a valine (V) residue; The consensus site 87 of the conserved insertion region is a leucine (L) residue; The consensus site 94 of the conserved insertion region is a glycine (G) residue; The consensus site 162 of the conserved insertion region is a glycine (G) residue; The consensus site 164 in the conserved insertion region is a tyrosine (Y) residue; The consensus site 165 of the conserved insertion region is a tryptophan (W) residue; The consensus site 184 in the conserved insertion region is an arginine (R) residue; The consensus site 197 in the conserved insertion region is a glycine (G) residue; The consensus site 204 of the conserved insertion region is a phenylalanine (F) residue; The consensus site 232 in the conserved insertion region is a leucine (L) residue; The consensus site 241 of the conserved insertion region is a phenylalanine (F) residue; The consensus site 243 in the conserved insertion region is a lysine (K) residue; The consensus site 249 in the conserved insertion region is a proline (P) residue; The consensus site 258 in the conserved insertion region is a leucine (L) residue; The consensus site 273 of the conserved insertion region is an aspartic acid (D) residue; The consensus site 280 in the conserved insertion region is a glutamine (Q) residue; The consensus site 291 of the conserved insertion region is a tyrosine (Y) residue; The consensus site 316 of the conserved insertion region is a proline (P) residue; The consensus site 327 of the conserved insertion region is a glycine (G) residue; and / or The consensus site 343 of the conserved insertion region is a proline (P) residue.
[0098] In some embodiments, the amino acid sequence of the CRISPR nuclease comprises a conserved insertion consisting of 343 consensus sites, each of which may be an amino acid or gap, such that removing the gap and joining amino acids results in the entire conserved insertion sequence; Consensus site 1 of the conserved insertion region is an asparagine (N) residue; Consensus site 2 of the conserved insertion region is a lysine (K), asparagine (N), or serine (S) residue; Consensus site 3 of the conserved insertion region is an isoleucine (I), methionine (M), or valine (V) residue; Consensus site 4 of the conserved insertion region is an isoleucine (I), leucine (L), or valine (V) residue; Consensus site 8 of the conserved insertion region is a proline (P), serine (S), or tyrosine (Y) residue; The consensus site 19 of the conserved insertion region is a cysteine (C), isoleucine (I), or leucine (L) residue, or a gap; The consensus site 29 of the conserved insertion region is an isoleucine (I), valine (V), tryptophan (W), or tyrosine (Y) residue; The consensus site 50 of the conserved insertion region is a proline (P) residue; The consensus site 56 of the conserved insertion region is an aspartic acid (D), glutamic acid (E), or glutamine (Q) residue; The consensus site 65 of the conserved insertion region is a glutamic acid (E), glycine (G), or asparagine (N) residue, or a gap; The consensus site 66 of the conserved insertion region is a glycine (G), histidine (H), lysine (K), asparagine (N), or threonine (T) residue; The consensus site 76 of the conserved insertion region is a glutamic acid (E) or arginine (R) residue, or a gap; The consensus site 77 of the conserved insertion region is a histidine (H) or leucine (L) residue, or a gap; The consensus site 87 of the conserved insertion region is a glutamic acid (E), leucine (L), or arginine (R) residue; The consensus site 89 of the conserved insertion region is a glycine (G), leucine (L), or glutamine (Q) residue, or a gap; The consensus site 90 of the conserved insertion region is an aspartic acid (D), lysine (K), asparagine (N), or serine (S) residue, or a gap; The consensus site 91 of the conserved insertion region is a glycine (G), asparagine (N), or tyrosine (Y) residue, or a gap; The consensus site 92 of the conserved insertion region is a glutamic acid (E), phenylalanine (F), or valine (V) residue, or a gap; The consensus site 94 of the conserved insertion region is a glycine (G) residue; The consensus site 96 of the conserved insertion region is an alanine (A), histidine (H), or tyrosine (Y) residue; The consensus site 100 of the conserved insertion region is a leucine (L), valine (V), or tyrosine (Y) residue; The consensus site 105 of the conserved insertion region is a glycine (G), lysine (K), or asparagine (N) residue, or a gap; The consensus site 113 of the conserved insertion region is a phenylalanine (F), histidine (H), lysine (K), or tyrosine (Y) residue; The consensus site 117 of the conserved insertion region is an alanine (A) or proline (P) residue; The consensus site 125 of the conserved insertion region is an isoleucine (I), arginine (R), or valine (V) residue; The consensus site 128 of the conserved insertion region is an alanine (A), glycine (G), threonine (T), or valine (V) residue; The consensus site 132 of the conserved insertion region is a leucine (L) residue or a gap; The consensus site 133 of the conserved insertion region is an asparagine (N) residue or a gap; The consensus site 134 of the conserved insertion region is a lysine (K) or arginine (R) residue, or a gap; The consensus site 135 of the conserved insertion region is an aspartic acid (D) residue or a gap; The consensus site 136 of the conserved insertion region is a glycine (G) or serine (S) residue, or a gap; The consensus site 140 of the conserved insertion region is a phenylalanine (F) or leucine (L) residue; The consensus site 147 of the conserved insertion region is a phenylalanine (F) residue or a gap; The consensus site 156 of the conserved insertion region is an aspartic acid (D), arginine (R), or serine (S) residue, or a gap; The consensus site 162 of the conserved insertion region is a glycine (G) or serine (S) residue; The consensus site 164 of the conserved insertion region is a cysteine (C) or tyrosine (Y) residue; The consensus site 165 of the conserved insertion region is a lysine (K) or tryptophan (W) residue; The consensus site 176 of the conserved insertion region is a valine (V) or tryptophan (W) residue, or a gap; The consensus site 186 of the conserved insertion region is an alanine (A), lysine (K), or arginine (R) residue, or a gap; The consensus site 187 of the conserved insertion region is a threonine (T) residue or a gap; The consensus site 195 of the conserved insertion region is a leucine (L) or valine (V) residue; The consensus site 197 in the conserved insertion region is a glycine (G) residue; The consensus site 199 of the conserved insertion region is an isoleucine (I), leucine (L), threonine (T), or valine (V) residue; The consensus site 204 of the conserved insertion region is a phenylalanine (F) residue; The consensus site 208 of the conserved insertion region is a proline (P) residue or a gap; The consensus site 209 of the conserved insertion region is an aspartic acid (D) residue or a gap; The consensus site 211 of the conserved insertion region is a phenylalanine (F) or tyrosine (Y) residue; The consensus site 215 of the conserved insertion region is a glutamic acid (E) or valine (V) residue, or a gap; The consensus site 216 of the conserved insertion region is an aspartic acid (D) or glutamic acid (E) residue, or a gap; The consensus site 217 of the conserved insertion region is an asparagine (N) or serine (S) residue, or a gap; The consensus site 218 of the conserved insertion region is a serine (S) residue or a gap; The consensus site 219 of the conserved insertion region is a valine (V) residue or a gap; The consensus site 225 of the conserved insertion region is an aspartic acid (D), histidine (H), or proline (P) residue; The consensus site 226 of the conserved insertion region is a glycine (G) or arginine (R) residue; The consensus site 235 of the conserved insertion region is an aspartic acid (D), glutamic acid (E), phenylalanine (F), methionine (M), or asparagine (N) residue; The consensus site 241 of the conserved insertion region is a phenylalanine (F), isoleucine (I), or tyrosine (Y) residue; The consensus site 249 of the conserved insertion region is a glycine (G), proline (P), or serine (S) residue; The consensus site 260 of the conserved insertion region is an alanine (A) or glycine (G) residue; The consensus site 268 of the conserved insertion region is an alanine (A), phenylalanine (F), or leucine (L) residue; The consensus site 273 of the conserved insertion region is an aspartic acid (D) or asparagine (N) residue; The consensus site 274 of the conserved insertion region is a glycine (G) residue or a gap; The consensus site 288 of the conserved insertion region is a glycine (G), serine (S), or threonine (T) residue; The consensus site 289 of the conserved insertion region is a lysine (K), glutamine (Q), or arginine (R) residue; The consensus site 290 of the conserved insertion region is a histidine (H) or tyrosine (Y) residue; The consensus site 303 of the conserved insertion region is a glutamic acid (E) or lysine (K) residue, or a gap; The consensus site 304 of the conserved insertion region is a glutamic acid (E) residue or a gap; The consensus site 305 of the conserved insertion region is a glutamine (Q) or arginine (R) residue, or a gap; The consensus site 306 of the conserved insertion region is a valine (V) residue or a gap; The consensus site 307 of the conserved insertion region is a threonine (T) residue or a gap; The consensus site 316 of the conserved insertion region is a proline (P) or glutamine (Q) residue; The consensus site 326 of the conserved insertion region is an aspartic acid (D) or glutamic acid (E) residue; The consensus site 327 of the conserved insertion region is a glycine (G) residue; The consensus site 331 of the conserved insertion region is an alanine (A) or leucine (L) residue, or a gap; The consensus site 333 of the conserved insertion region is an aspartic acid (D) or asparagine (N) residue; The consensus site 338 of the conserved insertion region is a glutamic acid (E), leucine (L), or arginine (R) residue; The consensus site 339 of the conserved insertion region is an alanine (A), isoleucine (I), or tyrosine (Y) residue; The consensus site 341 of the conserved insertion region is a phenylalanine (F) or leucine (L) residue; and / or The consensus site 343 of the conserved insertion region is a proline (P) or serine (S) residue.
[0099] In some embodiments, the amino acid sequence of the CRISPR nuclease is other than SEQ ID NOs: 45-48.
[0100] In some embodiments, the CRISPR nuclease is a nickase having an inactivated RuvC domain created by amino acid substitutions in the CRISPR nuclease at the positions shown in column 5 of Table 1.
[0101] In some embodiments, the CRISPR nuclease is a nickase having an inactive HNH domain created by amino acid substitutions in the CRISPR nuclease at the positions shown in column 6 of Table 1.
[0102] In some embodiments, the CRISPR nuclease is a dead nuclease having an inactivated RuvC domain and an inactivated HNH domain created by substitution of the CRISPR nuclease at the positions shown in column 7 of Table 1.
[0103] For example, substituting the aspartic acid residue (D) at position 7 of the OMNI-XL-12 amino acid sequence (SEQ ID NO: 12) with another amino acid (e.g., alanine (A)) can inactivate the RuvC domain and create a nickase of the OMNI-XL-12 nuclease. Substitutions with other amino acids are permitted at each of the amino acid positions listed in columns 5-7 unless otherwise indicated in Table 1. Other nickases or catalytically inactive nucleases can also be generated using the same designations as in Table 1.
[0104] In some embodiments, the CRISPR nuclease utilizes a protospacer adjacent motif (PAM) sequence shown in columns 2-3 of Table 3.
[0105] The invention also provides methods for modifying the nucleotide sequence of a DNA target site in the genome of a cell-free system or a cell, comprising introducing into the cell any one of the above compositions. In some embodiments, the composition contains a CRISPR nuclease and a crRNA:tracrRNA complex or an sgRNA molecule.
[0106] In some embodiments, the CRISPR nuclease cleaves the DNA strand adjacent to the CRISPR nuclease protospacer adjacent motif (PAM) sequence shown in columns 2-3 of Table 3, and cleaves the DNA strand adjacent to the sequence complementary to the PAM sequence. For example, an OMNI-XL-12 nuclease with an appropriate targeting sgRNA or crRNA:tracrRNA complex can cleave DNA at the strand adjacent to the sequence NRNNCCNN or NRTNCCRN, and at the DNA strand adjacent to the sequence complementary to the sequence NRNNCCNN or NRTNCCRN. In some embodiments, the DNA strand is within the nucleus of a cell.
[0107] In some embodiments, the CRISPR nuclease is a nickase having an inactivated RuvC domain created by amino acid substitutions in the CRISPR nuclease at the positions shown in column 5 of Table 1, and cleaves the DNA strand adjacent to the sequence complementary to the PAM sequence.
[0108] In some embodiments, the CRISPR nuclease is a nickase with an inactive HNH domain created by an amino acid substitution in the CRISPR nuclease at the position shown in column 6 of Table 1, and cleaves the DNA strand adjacent to the PAM sequence.
[0109] In some embodiments, the CRISPR nuclease is an inactivated nuclease having an inactivated RuvC domain and an inactivated HNH domain created by substitution of the CRISPR nuclease at the positions shown in column 7 of Table 1, and cleaves the DNA strand adjacent to the PAM sequence.
[0110] In some aspects, the cell is a eukaryotic cell or a prokaryotic cell.
[0111] In some aspects, the cells are mammalian cells.
[0112] In some embodiments, the cells are human cells.
[0113] In some embodiments, the CRISPR nuclease comprises an amino acid sequence having at least 100%, 99%, 98%, 97%, 96%, 95%, 94%, 93%, 92%, 91%, 90%, 89%, 88%, 87%, 86%, 85%, 84%, 83%, or 82% amino acid sequence identity to a CRISPR nuclease set forth in any of SEQ ID NOs: 1-22. In certain embodiments, the sequence encoding the CRISPR nuclease has at least 95% identity to a sequence selected from the group consisting of the nucleic acid sequences set forth in SEQ ID NOs: 23-44.
[0114] In some aspects of the invention, the disclosed compositions contain a DNA construct or vector system comprising a nucleotide sequence encoding a CRISPR nuclease or a CRISPR nuclease variant. In some embodiments, the nucleotide sequence encoding the CRISPR nuclease or a CRISPR nuclease variant is operably linked to a promoter operable in a cell of interest. In some embodiments, the cell of interest is a eukaryotic cell. In some embodiments, the cell of interest is a mammalian cell. In some embodiments, the nucleic acid sequence encoding the modified CRISPR nuclease is codon-optimized for use in cells of a particular organism. In some embodiments, the nucleic acid sequence encoding the nuclease is codon-optimized for E. coli. In some embodiments, the nucleic acid sequence encoding the nuclease is codon-optimized for eukaryotic cells. In some embodiments, the nucleic acid sequence encoding the nuclease is codon-optimized for mammalian cells.
[0115] In some embodiments, the composition contains a recombinant nucleic acid comprising a heterologous promoter operably linked to a polynucleotide encoding a CRISPR enzyme having at least 100%, 99%, 98%, 97%, 96%, 95%, 94%, 93%, 92%, 91%, or 90% identity to any of SEQ ID NOS: 1-22. Each possibility is a separate embodiment.
[0116] In one aspect of the composition, the CRISPR nuclease has at least 75%, 80%, 85, 90%, 95% or 97% identity to the amino acid sequence set forth in SEQ ID NO: 1, or the sequence encoding the CRISPR nuclease has at least 75%, 80%, 85, 90%, 95% or 97% sequence identity to the nucleotide sequence set forth in SEQ ID NO: 23.
[0117] In one aspect of the composition, the CRISPR nuclease has at least 75%, 80%, 85, 90%, 95% or 97% identity to the amino acid sequence set forth in SEQ ID NO:2, or the sequence encoding the CRISPR nuclease has at least 75%, 80%, 85, 90%, 95% or 97% sequence identity to the nucleotide sequence set forth in SEQ ID NO:24.
[0118] In one aspect of the composition, the CRISPR nuclease has at least 75%, 80%, 85, 90%, 95% or 97% identity to the amino acid sequence set forth in SEQ ID NO:3, or the sequence encoding the CRISPR nuclease has at least 75%, 80%, 85, 90%, 95% or 97% sequence identity to the nucleotide sequence set forth in SEQ ID NO:25.
[0119] In one aspect of the composition, the CRISPR nuclease has at least 75%, 80%, 85, 90%, 95% or 97% identity to the amino acid sequence set forth in SEQ ID NO:4, or the sequence encoding the CRISPR nuclease has at least 75%, 80%, 85, 90%, 95% or 97% sequence identity to the nucleotide sequence set forth in SEQ ID NO:26.
[0120] In one aspect of the composition, the CRISPR nuclease has at least 75%, 80%, 85, 90%, 95% or 97% identity to the amino acid sequence set forth in SEQ ID NO:5, or the sequence encoding the CRISPR nuclease has at least 75%, 80%, 85, 90%, 95% or 97% sequence identity to the nucleotide sequence set forth in SEQ ID NO:27.
[0121] In one aspect of the composition, the CRISPR nuclease has at least 75%, 80%, 85, 90%, 95% or 97% identity to the amino acid sequence set forth in SEQ ID NO:6, or the sequence encoding the CRISPR nuclease has at least 75%, 80%, 85, 90%, 95% or 97% sequence identity to the nucleotide sequence set forth in SEQ ID NO:28.
[0122] In one aspect of the composition, the CRISPR nuclease has at least 75%, 80%, 85, 90%, 95% or 97% identity to the amino acid sequence set forth in SEQ ID NO:7, or the sequence encoding the CRISPR nuclease has at least 75%, 80%, 85, 90%, 95% or 97% sequence identity to the nucleotide sequence set forth in SEQ ID NO:29.
[0123] In one aspect of the composition, the CRISPR nuclease has at least 75%, 80%, 85, 90%, 95% or 97% identity to the amino acid sequence set forth in SEQ ID NO:8, or the sequence encoding the CRISPR nuclease has at least 75%, 80%, 85, 90%, 95% or 97% sequence identity to the nucleotide sequence set forth in SEQ ID NO:30.
[0124] In one aspect of the composition, the CRISPR nuclease has at least 75%, 80%, 85, 90%, 95% or 97% identity to the amino acid sequence set forth in SEQ ID NO:9, or the sequence encoding the CRISPR nuclease has at least 75%, 80%, 85, 90%, 95% or 97% sequence identity to the nucleotide sequence set forth in SEQ ID NO:31.
[0125] In one aspect of the composition, the CRISPR nuclease has at least 75%, 80%, 85, 90%, 95% or 97% identity to the amino acid sequence set forth in SEQ ID NO: 10, or the sequence encoding the CRISPR nuclease has at least 75%, 80%, 85, 90%, 95% or 97% sequence identity to the nucleotide sequence set forth in SEQ ID NO: 32.
[0126] In one aspect of the composition, the CRISPR nuclease has at least 75%, 80%, 85, 90%, 95% or 97% identity to the amino acid sequence set forth in SEQ ID NO: 11, or the sequence encoding the CRISPR nuclease has at least 75%, 80%, 85, 90%, 95% or 97% sequence identity to the nucleotide sequence set forth in SEQ ID NO: 33.
[0127] In one aspect of the composition, the CRISPR nuclease has at least 75%, 80%, 85, 90%, 95% or 97% identity to the amino acid sequence set forth in SEQ ID NO: 12, or the sequence encoding the CRISPR nuclease has at least 75%, 80%, 85, 90%, 95% or 97% sequence identity to a sequence selected from the group consisting of the nucleotide sequences set forth in SEQ ID NOs: 34 and 392.
[0128] In one aspect of the composition, the CRISPR nuclease has at least 75%, 80%, 85, 90%, 95% or 97% identity to the amino acid sequence set forth in SEQ ID NO: 13, or the sequence encoding the CRISPR nuclease has at least 75%, 80%, 85, 90%, 95% or 97% sequence identity to the nucleotide sequence set forth in SEQ ID NO: 35.
[0129] In one aspect of the composition, the CRISPR nuclease has at least 75%, 80%, 85, 90%, 95% or 97% identity to the amino acid sequence set forth in SEQ ID NO: 14, or the sequence encoding the CRISPR nuclease has at least 75%, 80%, 85, 90%, 95% or 97% sequence identity to the nucleotide sequence set forth in SEQ ID NO: 36.
[0130] In one aspect of the composition, the CRISPR nuclease has at least 75%, 80%, 85, 90%, 95% or 97% identity to the amino acid sequence set forth in SEQ ID NO: 15, or the sequence encoding the CRISPR nuclease has at least 75%, 80%, 85, 90%, 95% or 97% sequence identity to the nucleotide sequence set forth in SEQ ID NO: 37.
[0131] In one aspect of the composition, the CRISPR nuclease has at least 75%, 80%, 85, 90%, 95% or 97% identity to the amino acid sequence set forth in SEQ ID NO: 16, or the sequence encoding the CRISPR nuclease has at least 75%, 80%, 85, 90%, 95% or 97% sequence identity to a sequence selected from the group consisting of the nucleotide sequences set forth in SEQ ID NOs: 38 and 393.
[0132] In one aspect of the composition, the CRISPR nuclease has at least 75%, 80%, 85, 90%, 95% or 97% identity to the amino acid sequence set forth in SEQ ID NO: 17, or the sequence encoding the CRISPR nuclease has at least 75%, 80%, 85, 90%, 95% or 97% sequence identity to the nucleotide sequence set forth in SEQ ID NO: 39.
[0133] In one aspect of the composition, the CRISPR nuclease has at least 75%, 80%, 85, 90%, 95% or 97% identity to the amino acid sequence set forth in SEQ ID NO: 18, or the sequence encoding the CRISPR nuclease has at least 75%, 80%, 85, 90%, 95% or 97% sequence identity to a sequence selected from the group consisting of the nucleotide sequences set forth in SEQ ID NOs: 40 and 394.
[0134] In one aspect of the composition, the CRISPR nuclease has at least 75%, 80%, 85, 90%, 95% or 97% identity to the amino acid sequence set forth in SEQ ID NO: 19, or the sequence encoding the CRISPR nuclease has at least 75%, 80%, 85, 90%, 95% or 97% sequence identity to a sequence selected from the group consisting of the nucleotide sequences set forth in SEQ ID NOs: 41 and 395.
[0135] In one aspect of the composition, the CRISPR nuclease has at least 75%, 80%, 85, 90%, 95% or 97% identity to the amino acid sequence set forth in SEQ ID NO:20, or the sequence encoding the CRISPR nuclease has at least 75%, 80%, 85, 90%, 95% or 97% sequence identity to the nucleotide sequence set forth in SEQ ID NO:42.
[0136] In one aspect of the composition, the CRISPR nuclease has at least 75%, 80%, 85, 90%, 95% or 97% identity to the amino acid sequence set forth in SEQ ID NO: 21, or the sequence encoding the CRISPR nuclease has at least 75%, 80%, 85, 90%, 95% or 97% sequence identity to the nucleotide sequence set forth in SEQ ID NO: 43.
[0137] In one aspect of the composition, the CRISPR nuclease has at least 75%, 80%, 85, 90%, 95% or 97% identity to the amino acid sequence set forth in SEQ ID NO: 22, or the sequence encoding the CRISPR nuclease has at least 75%, 80%, 85, 90%, 95% or 97% sequence identity to the nucleotide sequence set forth in SEQ ID NO: 44.
[0138] In some embodiments, there is provided an engineered or non-naturally occurring composition comprising a CRISPR nuclease comprising a sequence having at least 100%, 99%, 98%, 97%, 96%, 95%, 94%, 93%, 92%, 91%, 90%, 85%, or 80% identity to an amino acid sequence selected from the group consisting of SEQ ID NOs: 1-22, or a nucleic acid molecule comprising a sequence encoding said CRISPR nuclease. Each possibility represents a separate embodiment.
[0139] In some embodiments, CRISPR nucleases are engineered or non-naturally occurring. CRISPR nucleases may be recombinant. Such CRISPR nucleases pool genetic material from multiple sources and use laboratory methods (e.g., molecular cloning) to create sequences not otherwise found in organisms.
[0140] In certain embodiments, the CRISPR nuclease further comprises an RNA binding site capable of interacting with a DNA-targeting RNA molecule (gRNA molecule), and an active site that exhibits site-specific enzymatic activity.
[0141] In certain embodiments, the composition further comprises a DNA-targeting RNA molecule or a DNA polynucleotide encoding a DNA-targeting RNA molecule, wherein the DNA-targeting RNA molecule comprises a guide sequence portion, i.e., a nucleotide sequence complementary to a sequence in the target region, and wherein the DNA-targeting RNA molecule and the CRISPR nuclease do not occur together in nature.
[0142] In certain embodiments, the DNA-targeting RNA molecule further comprises a nucleotide sequence that is capable of forming a complex with a CRISPR nuclease.
[0143] This invention is One or more RNA molecules comprising a guide sequence portion linked to direct repeats capable of hybridizing to a target sequence, or one or more nucleotide sequences encoding said one or more RNA molecules; and a CRISPR nuclease comprising an amino acid sequence having at least 95% identity to a sequence selected from the group consisting of the amino acid sequences set forth in SEQ ID NOs: 1 to 22, or a nucleic acid molecule comprising a sequence encoding the CRISPR nuclease; Also provided are non-naturally occurring compositions containing a CRISPR-associated system comprising: wherein said one or more RNA molecules hybridize to said target sequence; wherein the target sequence is 3' of a protospacer adjacent motif (PAM), and the one or more RNA molecules form a complex with an RNA-guided nuclease.
[0144] In some embodiments, the composition further comprises an RNA molecule (e.g., a tracrRNA molecule) comprising a nucleotide sequence capable of complexing with a CRISPR nuclease, or a DNA polynucleotide comprising a sequence encoding an RNA molecule capable of complexing with a CRISPR nuclease.
[0145] In some embodiments, the composition further comprises a donor template for homology-directed repair (HDR).
[0146] In some embodiments, the composition is capable of editing a target region of the genome of a cell.
[0147] According to some aspects, (a) an RNA binding site; and Active sites that exhibit site-specific enzymatic activity A CRISPR nuclease or a polynucleotide encoding the CRISPR nuclease, wherein the CRISPR nuclease has at least 100%, 99%, 98%, 97%, 96%, 95%, 94%, 93%, 92%, 91%, 90%, 85%, or 80% identity to a sequence set forth in any of SEQ ID NOs: 1 to 22; and (b) i) a DNA-targeting RNA sequence comprising a nucleotide sequence complementary to a sequence in a target DNA sequence; and ii) protein-binding RNA sequences that can interact with the RNA-binding site of the CRISPR nuclease; one or more RNA molecules or DNA polynucleotides encoding said one or more RNA molecules, comprising: A naturally occurring composition is provided, comprising: Here, the DNA-targeting RNA sequence and the CRISPR nuclease do not occur together in nature, and each possibility is a separate embodiment.
[0148] In some embodiments, a single RNA molecule is provided, comprising a DNA targeting RNA sequence and a protein-binding RNA sequence, wherein the RNA molecule can form a complex with CRISPR nuclease and function as a DNA targeting module.In some embodiments, the length of the RNA molecule is at most 1000 bases, 900 bases, 800 bases, 700 bases, 600 bases, 500 bases, 400 bases, 300 bases, 200 bases, 100 bases, 50 bases.Each possibility is a separate embodiment.In some embodiments, the first RNA molecule comprising the DNA targeting RNA sequence and the second RNA molecule comprising the protein-binding RNA sequence interact by base pairing or are fused with each other to form a complex with CRISPR nuclease and form one or more RNA molecules that function as a DNA targeting module.
[0149] This invention is A CRISPR nuclease comprising a sequence having at least 95% identity to a sequence selected from the group consisting of the amino acid sequences set forth in SEQ ID NOs: 1 to 22, or a nucleic acid molecule comprising a sequence encoding the CRISPR nuclease; and a nucleotide sequence of a nuclease-binding RNA capable of interacting with / binding to the CRISPR nuclease; and a nucleotide sequence of the DNA-targeting RNA that comprises a sequence complementary to a sequence in the target DNA sequence; one or more RNA molecules, or one or more DNA polynucleotides encoding said one or more RNA molecules, comprising at least one of: Also provided are non-naturally occurring compositions containing a CRISPR-associated system comprising: Here, the CRISPR nuclease can complex with the one or more RNA molecules to form a complex that can hybridize to the target DNA sequence.
[0150] In some embodiments, the CRISPR nuclease and one or more RNA molecules form a CRISPR complex that is capable of binding to and cleaving a target DNA sequence.
[0151] In some embodiments, the CRISPR nuclease and at least one of the one or more RNA molecules do not occur together in nature.
[0152] In one aspect, CRISPR nucleases contain an RNA-binding site and an active site that exhibits site-specific enzymatic activity; The nucleotide sequence of the DNA-targeting RNA comprises a nucleotide sequence complementary to a sequence in the target DNA sequence; and The nucleotide sequence of the nuclease-binding RNA comprises a sequence that interacts with the RNA-binding site of the CRISPR nuclease.
[0153] In one embodiment, the nucleotide sequence of the nuclease-binding RNA and the nucleotide sequence of the DNA-targeting RNA are on a single guide RNA molecule (sgRNA), wherein the sgRNA molecule is capable of forming a complex with a CRISPR nuclease and functioning as a DNA-targeting module.
[0154] In one aspect, the nucleotide sequence of the nuclease-binding RNA is on a first RNA molecule and the nucleotide sequence of the DNA-targeting RNA is on a second RNA molecule, and the first and second RNA molecules interact by base pairing or are fused to each other to form an RNA complex, or sgRNA, that complexes with the CRISPR nuclease and functions as the DNA-targeting module.
[0155] In certain embodiments, the sgRNA is up to 1000 bases, 900 bases, 800 bases, 700 bases, 600 bases, 500 bases, 400 bases, 300 bases, 200 bases, 100 bases, or 50 bases in length.
[0156] In some embodiments, the composition further comprises a donor template for homology-directed repair (HDR).
[0157] In some embodiments, the CRISPR nuclease is not naturally occurring.
[0158] In some embodiments, the CRISPR nuclease is modified to include unnatural or synthetic amino acids.
[0159] In some embodiments, the CRISPR nuclease is modified to include one or more of a nuclear localization sequence (NLS), a cell-penetrating peptide sequence, and / or an affinity tag.
[0160] In certain embodiments, the CRISPR nuclease comprises one or more nuclear localization sequences of sufficient strength to drive accumulation of a detectable amount of a CRISPR complex comprising the CRISPR nuclease in the nucleus of a eukaryotic cell.
[0161] The present invention also provides a method for modifying the nucleotide sequence of a target site in the genome of a cell-free system or a cell, comprising introducing a composition of the present invention into the cell.
[0162] In some embodiments, the cell is a eukaryotic cell.
[0163] In another embodiment, the cell is a prokaryotic cell.
[0164] In some embodiments, the one or more RNA molecules further comprise an RNA sequence comprising a nucleotide molecule capable of complexing with an RNA nuclease (tracrRNA) or a DNA polynucleotide encoding an RNA molecule comprising a nucleotide sequence capable of complexing with a CRISPR nuclease.
[0165] In some embodiments, the CRISPR nuclease comprises one, two, three, four, five, six, seven, eight, nine, ten, or more NLSs at or near the amino terminus, one, two, three, four, five, six, seven, eight, nine, ten, or more NLSs at or near the carboxyl terminus, or a combination of one, two, three, four, five, six, seven, eight, nine, ten, or more NLSs at or near the amino terminus and one, two, three, four, five, six, seven, eight, nine, ten, or more NLSs at or near the carboxyl terminus. In some embodiments, one to four NLSs are fused to the CRISPR nuclease. In some embodiments, the NLS is internal to the open reading frame (ORF) of the CRISPR nuclease.
[0166] The method of fusing NLS at or near the amino terminus, at or near the carboxyl terminus, or within ORF of expressed protein is widely known in the art.For example, to fuse NLS to the amino terminus of CRISPR nuclease, the nucleic acid sequence of NLS is placed immediately after the start codon of CRISPR nuclease in the nucleic acid encoding NLS-fused CRISPR nuclease.Furthermore, to fuse NLS to the carboxyl terminus of CRISPR nuclease, the nucleic acid sequence of NLS is placed after the codon that codes the last amino acid of CRISPR nuclease and before the stop codon.
[0167] The present invention contemplates the combination of NLS, cell-penetrating peptide sequences and / or affinity tags positioned along the ORF of the CRISPR nuclease.
[0168] The amino acid and nucleic acid sequences of the CRISPR nucleases of the present invention may include NLSs and / or TAGs inserted to interrupt the consecutive amino acid or nucleic acid sequences of the CRISPR nuclease.
[0169] In certain embodiments, one or more NLSs are tandemly repeated.
[0170] In certain aspects, one or more NLSs are considered to be proximal to the N-terminus or C-terminus if the nearest amino acid of the NLS is within about 1, 2, 3, 4, 5, 10, 15, 20, 25, 30, 40, 50 or more amino acids along the polypeptide chain from the N-terminus or C-terminus.
[0171] As discussed, CRISPR nucleases may be modified to include one or more of a nuclear localization sequence (NLS), a cell-penetrating peptide sequence, and / or an affinity tag.
[0172] In certain embodiments, the CRISPR nuclease exhibits increased specificity for a target site compared to a wild-type CRISPR nuclease when complexed with one or more RNA molecules.
[0173] In certain embodiments, the complex of the CRISPR nuclease and one or more RNA molecules exhibits at least maintained on-target editing activity at the target site and reduced off-target activity compared to a wild-type CRISPR nuclease.
[0174] In certain embodiments, the composition further contains a recombinant nucleic acid molecule comprising a heterologous promoter operably linked to a nucleic acid molecule comprising a sequence encoding a CRISPR nuclease.
[0175] In certain embodiments, the CRISPR nuclease or the nucleic acid molecule comprising a sequence encoding the CRISPR nuclease is non-naturally occurring or modified.
[0176] The invention also provides non-naturally occurring or modified compositions containing vector systems comprising nucleic acid molecules comprising sequences encoding the CRISPR nucleases of the invention.
[0177] The invention also provides the use of the compositions of the invention in treating a subject suffering from a disease associated with a genomic mutation, comprising modifying a nucleotide sequence at a target site in the subject's genome.
[0178] The present invention provides a method for modifying a nucleotide sequence at a target site in the genome of a mammalian cell, comprising: (i) a composition containing a nucleic acid molecule comprising a sequence encoding a CRISPR nuclease having at least 95% identity to a sequence selected from the group consisting of the amino acid sequences set forth in SEQ ID NOs: 1 to 22, or a CRISPR nuclease having a sequence at least 95% identity to a sequence selected from the group consisting of the nucleic acid sequences set forth in SEQ ID NOs: 23 to 44; and (ii) a DNA-targeting RNA molecule comprising a nucleotide sequence complementary to the sequence of the target DNA, or a DNA polynucleotide encoding the DNA-targeting RNA molecule, into the cell.
[0179] In some embodiments, the method is performed ex vivo. In some embodiments, the method is performed in vivo. In some embodiments, some steps of the method are performed ex vivo and some steps are performed in vivo. In some embodiments, the mammalian cells are human cells.
[0180] In some embodiments, the method further comprises (iii) introducing into the cell an RNA molecule comprising the tracrRNA sequence or a DNA polynucleotide encoding an RNA molecule comprising the tracrRNA sequence.
[0181] In some embodiments, the DNA-targeting RNA molecule comprises a crRNA repeat sequence.
[0182] In some embodiments, an RNA molecule comprising a tracrRNA sequence can bind to a DNA-targeting RNA molecule.
[0183] In some embodiments, the DNA-targeting RNA molecule and the RNA molecule comprising the tracrRNA sequence interact to form an RNA complex, which can form an active complex with a CRISPR nuclease.
[0184] In one embodiment, the DNA-targeting RNA molecule and the RNA molecule comprising the nuclease-binding RNA sequence are fused together in the form of a single guide RNA molecule suitable for forming an active complex with a CRISPR nuclease.
[0185] In some embodiments, the guide sequence portion comprises a sequence complementary to a protospacer sequence.
[0186] In some embodiments, the CRISPR nuclease complexes with the DNA-targeting RNA molecule and makes a double-stranded break in the region 3' or 5' of the protospacer adjacent motif (PAM).
[0187] In some aspects of the methods described herein, the methods are for treating a subject suffering from a disease associated with a genomic mutation, comprising modifying a nucleotide sequence at a target site in the subject's genome.
[0188] In one aspect, the method involves first selecting a subject suffering from a disease associated with a genomic mutation and obtaining cells from the subject.
[0189] The invention also provides modified cells obtained by the methods described herein. In some embodiments, these modified cells are capable of giving rise to progeny cells. In some embodiments, these modified cells are capable of giving rise to progeny cells after transplantation.
[0190] The invention also provides compositions containing these modified cells and a pharmaceutically acceptable carrier, as well as in vitro or ex vivo methods for preparing the same, which involve combining the cells with a pharmaceutically acceptable carrier.
[0191] DNA-targeting RNA molecules The "guide sequence portion" of an RNA molecule refers to a nucleotide sequence that can hybridize to a specific target DNA sequence. For example, the guide sequence portion has a nucleotide sequence that is partially or completely complementary to the target DNA sequence along the guide sequence portion. In some embodiments, the nucleotide length of the guide sequence portion is 17, 18, 19, 20, 21, 22, 23, 24, 25, 26, 27, 28, 29, 30, 31, 32, 33, 34, 35, 36, 37, 38, 39, 40, 41, 42, 43, 44, 45, 46, 47, 48, 49, or 50 nucleotides, or about 17-50, 17-49, 17-48, 17-47, 17-46, 17-45, 17-44, 17-43, 17-42, 17-41, 17-44, 17-45, 17-46, 17-47, 17-48, 17-49, 17-49, 17-50, 17-51, 17-52, 17-53, 17-54, 17-55, 17-56, 17-57, 17-58, 17-59, 17-60, 17-61, 17-62, 17-63, 17-64, 17-65, 17-66, 17-67, 17-68, 17-69, 17-70, 17-71, 17-72, 17-73, 17-74, 17- 0, 17-39, 17-38, 17-37, 17-36, 17-35, 17-34, 17-33, 17-31, 17-30, 17-29, 17-28, 17-27, 17-26, 17-25, 17-24, 17-22, 17-21, 18-25, 18-24, 18-23, 18-22, 18-21, 19-25, 19-24, 19-23, 19-22, 19-21, 19-20, 20-22, 18-20, 20-21, 21-22, or 17-20. The entire length of the guide sequence portion is completely complementary to the target DNA sequence. The guide sequence portion may be a portion of an RNA molecule that can form a complex with a CRISPR nuclease, with the guide sequence portion serving as the DNA targeting portion of the CRISPR complex. When a DNA molecule having a guide sequence portion is present simultaneously with a CRISPR molecule, the RNA molecule can direct the CRISPR nuclease to a specific target DNA sequence. Each possibility is a separate embodiment. The RNA molecule can be specifically designed to target a desired sequence. Thus, a molecule containing a "guide sequence portion" is a type of targeting molecule. Throughout this application, the terms "guide molecule," "RNA guide molecule," "guide RNA molecule," and "gRNA molecule" are synonymous with a molecule containing a guide sequence portion, and the term "spacer" is synonymous with "guide sequence portion."
[0192] In aspects of the invention, CRISPR nucleases have greatest cleavage activity when used with RNA molecules that include a guide sequence portion that is 17, 18, 19, 20, 21, 22, 23, 24, 25, 26, 27, 28, 29, or 30 nucleotides long.
[0193] Single guide RNA (sgRNA) molecules can be used to direct CRISPR nucleases to desired target sites. Single guide RNAs include a guide sequence portion and a scaffold portion. The scaffold portion interacts with CRISPR nucleases, and together with the guide sequence portion, activates and directs CRISPR nucleases to desired target sites. The scaffold portion can be further modified, for example, to reduce its size.
[0194] In some aspects of the invention, the disclosed methods include methods for modifying the nucleotide sequence of a target site in the genome of a cell-free system or a cell, comprising introducing into the cell a composition of the embodiments described herein.
[0195] In some aspects, the cell is a eukaryotic cell, preferably a mammalian cell or a plant cell.
[0196] In some aspects of the invention, the disclosed methods also provide for the use of the compositions described herein in treating a subject suffering from a disease associated with a genomic mutation, comprising modifying a nucleotide sequence at a target site in the subject's genome.
[0197] In some aspects of the invention, the disclosed methods include methods of treating a subject with a mutational disorder, comprising targeting a composition described herein to an allele associated with the mutational disorder.
[0198] In some aspects, the mutational disorder is associated with a disease or disorder selected from neoplasia, age-related macular degeneration, schizophrenia, neurological disorders, neurodegenerative diseases, movement disorders, fragile X syndrome, secretase-related disorders, prion-related disorders, ALS, addiction, autism, Alzheimer's disease, neutropenia, inflammation-related disorders, Parkinson's disease, blood and coagulation diseases and disorders, beta thalassemia, sickle cell anemia, cellular dysregulation, tumor-related diseases and disorders, inflammation and immune-related diseases and disorders, metabolic, liver, hypercholesterolemia, kidney and protein diseases and disorders, muscular and skeletal diseases and disorders, skin diseases and disorders, neurological diseases and disorders, pulmonary diseases and disorders, corneal diseases and disorders, retinal diseases and disorders, and ophthalmic diseases and disorders.
[0199] CRISPR nuclease domain The characteristic targeted nuclease activity of CRISPR nucleases is conferred by the various functions of their specific domains. In this application, an OMNI-XL domain is defined as Domain A, Domain B, Domain C, Domain D, Domain E, Domain F, Domain G, Domain H, Domain I, Domain J, Domain K, Domain L, or Domain M.
[0200] In one aspect, Domain A is identified as amino acids 1-46 of SEQ ID NO:12, as amino acids 1-49 of SEQ ID NO:16, as amino acids 1-52 of SEQ ID NO:18, and as amino acids 1-52 of SEQ ID NO:19.
[0201] In one aspect, Domain B is identified as amino acids 47-89 of SEQ ID NO:12, as amino acids 50-84 of SEQ ID NO:16, as amino acids 53-86 of SEQ ID NO:18, and as amino acids 53-86 of SEQ ID NO:19.
[0202] In one aspect, Domain C is identified as amino acids 90 to 285 of SEQ ID NO:12, as amino acids 85 to 340 of SEQ ID NO:16, as amino acids 87 to 339 of SEQ ID NO:18, and as amino acids 87 to 339 of SEQ ID NO:19.
[0203] In one aspect, Domain D is identified as amino acids 286 to 545 of SEQ ID NO:12, as amino acids 341 to 596 of SEQ ID NO:16, as amino acids 340 to 603 of SEQ ID NO:18, and as amino acids 340 to 603 of SEQ ID NO:19.
[0204] In one aspect, Domain E is identified as amino acids 546 to 622 of SEQ ID NO:12, as amino acids 597 to 674 of SEQ ID NO:16, as amino acids 604 to 681 of SEQ ID NO:18, and as amino acids 604 to 681 of SEQ ID NO:19.
[0205] In one aspect, Domain F is identified as amino acids 623-653 of SEQ ID NO:12, as amino acids 675-705 of SEQ ID NO:16, as amino acids 682-712 of SEQ ID NO:18, and as amino acids 682-712 of SEQ ID NO:19.
[0206] In one aspect, Domain G is identified as amino acids 654 to 774 of SEQ ID NO:12, as amino acids 706 to 825 of SEQ ID NO:16, as amino acids 713 to 834 of SEQ ID NO:18, and as amino acids 713 to 834 of SEQ ID NO:19.
[0207] In one aspect, Domain H is identified as amino acids 775-789 of SEQ ID NO:12, as amino acids 826-838 of SEQ ID NO:16, as amino acids 835-845 of SEQ ID NO:18, and as amino acids 835-845 of SEQ ID NO:19.
[0208] In one aspect, Domain I is identified as amino acids 790-828 of SEQ ID NO:12, as amino acids 839-877 of SEQ ID NO:16, as amino acids 846-886 of SEQ ID NO:18, and as amino acids 846-886 of SEQ ID NO:19.
[0209] In one aspect, Domain J is identified as amino acids 829 to 1153 of SEQ ID NO:12, as amino acids 878 to 1179 of SEQ ID NO:16, as amino acids 887 to 1203 of SEQ ID NO:18, and as amino acids 887 to 1203 of SEQ ID NO:19.
[0210] In one aspect, Domain K is identified as amino acids 1154-1264 of SEQ ID NO:12, as amino acids 1180-1278 of SEQ ID NO:16, as amino acids 1204-1313 of SEQ ID NO:18, and as amino acids 1204-1313 of SEQ ID NO:19.
[0211] In one aspect, Domain L is identified as amino acids 1265-1360 of SEQ ID NO:12, as amino acids 1279-1372 of SEQ ID NO:16, as amino acids 1314-1411 of SEQ ID NO:18, and as amino acids 1314-1407 of SEQ ID NO:19.
[0212] In one aspect, Domain M is identified as amino acids 1361 to 1511 of SEQ ID NO:12, as amino acids 1373 to 1517 of SEQ ID NO:16, as amino acids 1412 to 1565 of SEQ ID NO:18, and as amino acids 1408 to 1566 of SEQ ID NO:19.
[0213] The extent of each domain may vary slightly depending on the parameters of the sequence alignment analysis, for example, the beginning or end of the range may vary by up to 10 amino acids.
[0214] The activity of each domain provides an aspect of each nuclease's advantageous characteristics. Descriptions of several CRISPR nuclease domains and their general functions can be found in, among other places, Mir et al., ACS Chem. Biol. (2019), Palermo et al., Quarterly Reviews of Biophysics (2018), Jiang and Doudna, Annual Review of Biophysics (2017), Nishimasu et al., Cell (2014), and Nishimasu et al., Cell (2015), which are incorporated herein by reference.
[0215] In one aspect of the invention, amino acid sequences having similarity to the OMNI-XL domain may be used to design and produce non-naturally occurring peptides (e.g., CRISPR nucleases) such that the peptides exhibit advantageous characteristics of the activity of the OMNI-XL domain.
[0216] In certain embodiments, such a peptide (e.g., a CRISPR nuclease) comprises an amino acid sequence having at least 100%, 99.5%, 99%, 98%, 97%, 96%, 95%, 94%, 93%, 92%, 91%, 90%, 89%, 88%, 87%, 86%, 85%, 84%, 83%, 82%, 81%, 80%, 79%, 78%, 77%, 76%, 75%, 74%, 73%, 72%, 71%, or 70% identity to the amino acid sequence of at least one of Domain A, Domain B, Domain C, Domain D, Domain E, Domain F, Domain G, Domain H, Domain I, Domain J, Domain K, Domain L, or Domain M of a nuclease. In some embodiments, identity is to at least one amino acid sequence, at least two amino acid sequences, at least three amino acid sequences, at least four amino acid sequences, at least five amino acid sequences, at least six amino acid sequences, at least seven amino acid sequences, or at least eight amino acid sequences of Domain A, Domain B, Domain C, Domain D, Domain E, Domain F, Domain G, Domain H, Domain I, Domain J, Domain K, Domain L, or Domain M of the nuclease. Each possibility is a separate embodiment. In some embodiments, identity is to at least one amino acid sequence of Domain A, Domain B, Domain C, Domain D, Domain E, Domain F, Domain G, Domain H, Domain I, Domain J, Domain K, Domain L, or Domain M of the nuclease. In some embodiments, identity is to the amino acid sequence of Domain J of the nuclease.In certain embodiments, peptides exhibit extensive amino acid variability relative to the full-length OMNI-XL amino acid sequence beyond the amino acid sequence of a peptide having at least 100%, 99.5%, 99%, 98%, 97%, 96%, 95%, 94%, 93%, 92%, 91%, 90%, 89%, 88%, 87%, 86%, 85%, 84%, 83%, 82%, 81%, 80%, 79%, 78%, 77%, 76%, 75%, 74%, 73%, 72%, 71%, or 70% identity to the amino acid sequence of at least one of Domain A, Domain B, Domain C, Domain D, Domain E, Domain F, Domain G, Domain H, Domain I, Domain J, Domain K, Domain L, or Domain M of OMNI-XL nuclease. In certain embodiments, a peptide comprises an amino acid sequence between two domain sequences. In one embodiment, the length of the amino acid sequence present between the domains is 1 to 10, 10 to 20, 20 to 40, 40 to 50, or up to 100. In one embodiment, the sequence present between the domains is a linker sequence.
[0217] In one aspect of the invention, the amino acid sequence encoding any one of the domains of the OMNI-XL nuclease described in the peptide specification may be substituted with one or more amino acids compared to the sequence of the original OMNI-XL domain. The amino acid substitutions may be conservative, i.e., with an amino acid having similar chemical properties to the original amino acid. For example, a positively charged amino acid may be substituted with another positively charged amino acid, e.g., an arginine residue may be substituted with a lysine residue, or a polar amino acid may be substituted with a different polar amino acid. Conservative substitutions are more tolerated, and an amino acid sequence encoding any one of the domains of the OMNI-XL nuclease may contain up to 10% of such substitutions. The amino acid substitutions may also be radical, i.e., with an amino acid having different chemical properties from the original amino acid. For example, a positively charged amino acid may be substituted with a negatively charged amino acid, e.g., an arginine residue may be substituted with a glutamic acid residue, or a polar amino acid may be substituted with a nonpolar amino acid. The amino acid substitutions may be semi-conservative, or the amino acid substitutions may be made with any other amino acid. The substitution may alter the activity from the original OMNI-XL domain function, for example, reducing catalytic nuclease activity.
[0218] In some aspects of the invention, the disclosed compositions include non-naturally occurring compositions containing a CRISPR nuclease, wherein the CRISPR nuclease comprises an amino acid sequence that corresponds to the amino acid sequence of at least one of Domain A, Domain B, Domain C, Domain D, Domain E, Domain F, Domain G, Domain H, Domain I, Domain J, Domain K, Domain L, or Domain M of OMNI-XL nuclease. In some embodiments of the invention, the CRISPR nuclease comprises at least one, at least two, at least three, at least four, or at least five amino acid sequences, each of which corresponds to any one of the amino acid sequences Domain A, Domain B, Domain C, Domain D, Domain E, Domain F, Domain G, Domain H, Domain I, Domain J, Domain K, Domain L, or Domain M of OMNI-XL nuclease. Thus, a CRISPR nuclease may comprise a combination of amino acid sequences corresponding to Domain A, Domain B, Domain C, Domain D, Domain E, Domain F, Domain G, Domain H, Domain I, Domain J, Domain K, Domain L, or Domain M of OMNI-XL nuclease. In some embodiments, the amino acid sequence is at least 100-250, 250-500, 500-1000, 1000-1500, 1000-1700, or 1000-2000 amino acids in length.
[0219] Diseases and Treatments Certain embodiments of the invention target nucleases to specific genetic loci associated with a disease or disorder as a form of gene editing, treatment, or therapy. For example, the novel nucleases disclosed herein may be specifically targeted to pathogenic mutant alleles of a gene using specially designed guide RNA molecules to induce gene editing or knockout. Preferably, guide RNA molecules are designed by first considering the PAM requirements of the nuclease, which will depend on the system in which gene editing is to be performed, as described herein. For example, guide RNA molecules designed to target the OMNI-XL-12 nuclease to a target site are designed to include a spacer region complementary to the DNA strand of the DNA double-stranded region adjacent to the OMNI-XL-12 PAM sequence (e.g., "NRNNCCNN" or "NRTNCCRN"). The guide RNA molecule is preferably further designed to include a spacer region (i.e., the region of the guide RNA molecule complementary to the target allele) of sufficient, and preferably optimal, length to increase the specific activity of the nuclease and reduce off-target effects.
[0220] As a non-limiting example, a guide RNA molecule may be designed to target a nuclease to a specific region of a mutant allele, such as near the start codon, so that upon DNA damage by the nuclease, the non-homologous end joining (NHEJ) pathway is induced, resulting in the silencing of the mutant allele by introducing a frameshift mutation. This approach to designing a guide RNA molecule is particularly useful for altering the effect of a dominant-negative mutation, thereby treating a subject. As another non-limiting example, a guide RNA molecule may be designed to target a specific pathogenic mutation of a mutated allele, so that upon DNA damage by the nuclease, the homology-directed repair (HDR) pathway is induced, resulting in template-mediated correction of the mutant allele. This approach to designing a guide RNA molecule is particularly useful for altering the haploinsufficient effect of a mutant allele, thereby treating a subject.
[0221] Non-limiting examples of genes that may be targeted for modification to treat a disease or disorder are provided below. Disease-related genes and mutations that cause mutational disorders have been described in the literature. Such mutations allow for the design of DNA-targeting RNA molecules that direct CRISPR compositions to alleles of disease-related genes, where the CRISPR compositions cause DNA damage and induce DNA repair pathways to modify the alleles, thereby treating the mutational disorder.
[0222] Mutations in the ELANE gene are associated with neutropenia, and therefore, without limitation, aspects of the invention that target ELANE may be used in methods of treating subjects suffering from neutropenia.
[0223] CXCR4 is a co-receptor in human immunodeficiency virus type 1 (HIV-1) infection. Accordingly, without limitation, embodiments of the invention that target CXCR4 may be used in methods of treating a subject with HIV-1 or conferring resistance to HIV-1 infection in a subject.
[0224] Disruption of programmed cell death protein 1 (PD-1) promotes CAR-T cell killing of tumor cells, making PD-1 a potential target for cancer therapy. Accordingly, without limitation, embodiments of the invention that target PD-1 may be used in methods of treating subjects with cancer. In one embodiment, the treatment is CAR-T cell therapy with T cells engineered according to the invention to be PD-1 deficient.
[0225] Furthermore, BCL11A is a gene involved in the suppression of hemoglobin production. Inhibiting BCL11A increases globin production and may treat diseases such as thalassemia and sickle cell anemia. See, for example, International Publication No. 2017 / 077394, U.S. Patent Application Publication No. 2011 / 0182867; Humbert et al. Sci. Transl. Med. (2019), and Canver et al. Nature (2015). Thus, without limitation, embodiments of the invention targeting the BCL11A enhancer may be used in methods for treating subjects with β-thalassemia or sickle cell anemia.
[0226] Aspects of this invention that target disease-associated genes may be used to study, modify, or treat the diseases or disorders listed below in Table A or Table B. Indeed, any disease associated with a genetic locus may be studied, modified, or treated using the nucleases disclosed herein to target the appropriate disease-associated gene (e.g., those listed in U.S. Patent Application Publication No. 2018 / 0282762 and EP 3079726 B1).
[0227] [Table A]
[0228] [Table B-1]
[0229] [Table B-2]
[0230] [Table B-3]
[0231] Unless otherwise defined, all technical and / or scientific terms used herein have the same meaning as commonly understood by one of ordinary skill in the art to which this invention pertains. Although methods and materials similar or equivalent to those described in the specification can be used in the practice or testing of embodiments of the invention, representative methods and / or materials are described below. In case of conflict, the specification, including definitions, will control. Additionally, the materials, methods, and examples are illustrative only and are not intended to be necessarily limiting.
[0232] Unless otherwise stated in the discussion section, adjectives such as "substantially" and "about" modifying a state or characteristic of a feature of an embodiment of the invention are understood to mean that the state or characteristic is defined within a range acceptable for operation of the embodiment in its intended use. Unless otherwise indicated, the term "or" in the specification and claims is considered an inclusive "or" rather than an exclusive "or," indicating at least one or any combination of the associated items.
[0233] It should be understood that the term "a" or "an" in the specification refers to "one or more" of the listed components. It will be clear to one of ordinary skill in the art that the use of the singular includes the plural unless otherwise specified. Thus, the terms "a" and "at least one" have the same meaning in this application.
[0234] To better understand the present teachings and in no way limit the scope of the teachings, unless otherwise specified, all numbers indicating quantities, percentages, or ratios, and other numerical values used in the specification and claims should be understood to be modified in all instances by the term "about." Thus, unless indicated to the contrary, the numerical values set forth in the specification and claims are approximations that may vary depending on the desired properties sought to be obtained. At the very least, each numerical value should be construed in light of the number of significant digits and by applying ordinary rounding techniques.
[0235] Where numerical ranges are stated herein, it is understood that the invention contemplates every integer between the upper and lower limits, inclusive, unless otherwise stated.
[0236] In this specification and claims, the verbs "contain," "include," and "have," and each of their conjugations, are used to indicate that the object of the verb is not necessarily an exhaustive list of components, elements, or parts of the subject of the verb. Other terms used in this specification have meanings commonly known in the art.
[0237] The terms "polynucleotide," "nucleotide," "nucleotide sequence," "nucleic acid," and "oligonucleotide" are synonymous. They refer to a polymeric form of nucleotides of any length, either deoxyribonucleotides or ribonucleotides, or their analogs. Polynucleotides may have any three-dimensional structure and may perform any function, known or unknown. Non-limiting examples of polynucleotides include coding and non-coding regions of a gene or gene fragment, loci determined by linkage analysis, exons, introns, messenger RNA (mRNA), transfer RNA, ribosomal RNA, small interfering RNA (siRNA), short hairpin RNA (shRNA), microRNA (miRNA), ribozymes, cDNA, recombinant polynucleotides, branched polynucleotides, plasmids, vectors, isolated DNA sequences, isolated RNA sequences, nucleic acid probes, and primers. A polynucleotide may contain one or more modified nucleotides, such as methylated nucleotides or their analogs. Modifications to the nucleotide structure may occur before or after assembly of the polymer. The nucleotide sequence may be interrupted by non-nucleotide components. A polynucleotide may be further modified after polymerization, such as by conjugation with a labeling component.
[0238] The term "nucleotide analog" or "modified nucleotide" refers to a nucleotide that contains one or more of various chemical modifications (e.g., substitutions) in the nitrogenous base of the nucleoside (e.g., cytosine (C), thymine (T) or uracil (U), adenine (A) or guanine (G)), in the sugar moiety of the nucleoside (e.g., ribose, deoxyribose, modified ribose, modified deoxyribose, six-membered sugar analog, or open-ring sugar analog), or in the phosphate moiety. Each of the RNA sequences described herein may contain one or more nucleotide analogs.
[0239] In this specification, the following nucleotide identifiers are used to represent nucleotide bases:
[0240] [Table C]
[0241] As used herein, the term "targeting sequence" or "targeting molecule" refers to a nucleotide sequence capable of hybridizing with a specific target sequence or a molecule comprising such a nucleotide sequence; for example, a targeting sequence has a nucleotide sequence that is at least partially complementary to the sequence to be targeted. The targeting sequence or targeting molecule may be a portion of a targeting RNA molecule capable of forming a complex with a CRISPR nuclease, where the targeting sequence serves as the targeting portion of the CRISPR complex. When a molecule having a targeting sequence is present simultaneously with a CRISPR molecule, the RNA molecule can direct the CRISPR nuclease to a specific target sequence. Each possibility is a separate embodiment. The targeting RNA molecule can be specifically designed to target a desired sequence.
[0242] In this specification, the term "targeting" or "directing to the target" refers to the preferential hybridization of a targeting molecule or targeting sequence with a nucleic acid having a target nucleotide sequence. The term "targeting" or "directing to the target" encompasses variable hybridization efficiency, and thus, although the nucleic acid having the target nucleotide sequence is preferentially targeted, it is understood that in addition to on-target hybridization, unintended off-target hybridization may also occur. When an RNA molecule targets a sequence, it is understood that the complex of the RNA molecule and the CRISPR nuclease molecule targets that sequence for nuclease activity.
[0243] When targeting a DNA sequence present in multiple cells, it is understood that targeting encompasses the hybridization of the guide sequence portion of the RNA molecule with the sequence in one or more cells, and also encompasses the hybridization of the RNA molecule with the target sequence in not all cells in multiple cells.Therefore, when targeting a sequence in multiple cells, it is understood that the complex of the RNA molecule and CRISPR nuclease hybridizes with the target sequence in one or more cells, and it is also understood that it may hybridize with the target sequence in not all cells.Therefore, it is understood that the complex of the RNA molecule and CRISPR nuclease may hybridize with the target sequence in one or more cells and cause double-strand breaks, and may hybridize with the target sequence in not all cells and cause double-strand breaks.In this specification, the term "modified cell" refers to a cell in which double-strand breaks are made by the complex of the RNA molecule and CRISPR nuclease as a result of hybridization with the target sequence, i.e., on-target hybridization.
[0244] As used herein, the term "wild-type" is a term of art understood by those skilled in the art and refers to the typical form of a naturally occurring organism, strain, gene, or trait, as distinguished from a variant or mutant. Thus, as used herein, when an amino acid or nucleotide sequence refers to a wild-type sequence, a variant refers to a variant of that sequence, including, for example, a substitution, deletion, or addition. In embodiments of the invention, the modified CRISPR nuclease is a variant of a CRISPR nuclease that includes at least one amino acid modification (e.g., a substitution, deletion, and / or addition) relative to any of the CRISPR nucleases listed in Table 1.
[0245] The terms "non-natural," "non-naturally occurring," or "modified" are used interchangeably and refer to human modification. When used with reference to a nucleic acid molecule or polypeptide, this term may mean that the nucleic acid molecule or polypeptide is at least substantially free from at least one component with which it is naturally associated and found in nature.
[0246] As used herein, the term "amino acid" includes natural and / or unnatural or synthetic amino acids, including glycine, their D or L, optical isomers, and amino acid analogs and peptidomimetics.
[0247] As used herein, "genomic DNA" refers to linear and / or chromosomal DNA and / or plasmid or other extrachromosomal DNA sequences present in a cell or cells of interest. In some embodiments, the cells of interest are eukaryotic cells. In some embodiments, the cells of interest are prokaryotic cells. In some embodiments, the method generates a double-strand break (DSB) at a predetermined target site in the genomic DNA sequence, resulting in a mutation, addition, and / or deletion of the DNA sequence at the target site in the genome.
[0248] "Eukaryotic" cells include, but are not limited to, fungal cells (such as yeast), plant cells, animal cells, mammalian cells, and human cells.
[0249] As used herein, the term "nuclease" refers to an enzyme capable of cleaving phosphodiester bonds between nucleotide subunits of nucleic acids. Nucleases may be isolated or derived from natural sources. The natural source may be any organism. Alternatively, nucleases may be modified or synthetic proteins with phosphodiester bond cleavage activity.
[0250] As used herein, the term "PAM" refers to a nucleotide sequence in a target DNA that is located adjacent to the target DNA sequence and that is recognized by a CRISPR nuclease. The PAM sequence may vary depending on the nuclease.
[0251] As used herein, the term "mutational disorder" or "mutational disease" refers to a disorder or disease associated with a dysfunction of a gene caused by a mutation. A dysfunctional gene that manifests as a mutational disorder contains a mutation in at least one of its alleles and is referred to as a "disease-associated gene." The mutation may be present in any part of the disease-associated gene, for example, a regulatory portion, a coding portion, or a non-coding portion. The mutation may be a substitution, addition, or deletion mutation. Mutations in disease-associated genes may manifest as disorders or diseases depending on any mutation mechanism, such as recessive, dominant-negative, gain-of-function, loss-of-function, or mutations leading to haploinsufficiency of the gene product.
[0252] Those skilled in the art will appreciate that embodiments of this invention disclose RNA molecules that can form a complex with a nuclease (e.g., a CRISPR nuclease), such as by binding to a target genomic DNA sequence of interest adjacent to a protospacer adjacent motif (PAM). The nuclease then cleaves the target DNA, creating a double-stranded break within the protospacer.
[0253] In some embodiments of the invention, the CRISPR nuclease and the targeting molecule form a CRISPR complex that binds to and cleaves a target DNA sequence. The CRISPR nuclease may form a CRISPR complex that includes the CRISPR nuclease and an RNA molecule without a separate tracrRNA molecule. Alternatively, the CRISPR nuclease may form a CRISPR complex with the CRISPR nuclease, an RNA molecule, and a tracrRNA molecule.
[0254] The term "protein binding sequence" or "nuclease binding sequence" refers to a sequence that can bind to a CRISPR nuclease to form a CRISPR complex. Those skilled in the art will understand that a tracrRNA that can bind to a CRISPR nuclease to form a CRISPR complex contains a protein or nuclease binding sequence.
[0255] The "RNA-binding portion" of a CRISPR nuclease refers to the portion of the CRISPR nuclease that can bind to an RNA molecule to form a CRISPR complex, such as the nuclease-binding sequence of a tracrRNA molecule. The "active portion" of a CRISPR nuclease refers to the portion of the CRISPR nuclease that makes a double-stranded break in a DNA molecule, for example, when complexed with a DNA-targeting RNA molecule.
[0256] The RNA molecule may comprise a sequence sufficiently complementary to the tracrRNA molecule so as to hybridize to the tracrRNA through base pairing and promote the formation of a CRISPR complex (see U.S. Patent No. 8,906,616). In some embodiments of the invention, the RNA molecule may further comprise a portion having a tracr mate sequence.
[0257] In some embodiments of the invention, the targeting molecule may further comprise the sequence of a tracrRNA molecule. Such embodiments may be designed as a synthetic fusion of a guide portion of an RNA molecule (gRNA or crRNA) and a transactivating crRNA (tracrRNA), forming a single guide RNA (sgRNA) (see Jinek et al., Science (2012)). In some embodiments of the invention, a CRISPR complex may also be formed that utilizes another tracrRNA molecule and another RNA molecule that includes a guide sequence portion. In such embodiments, the tracrRNA molecule may hybridize to the RNA molecule through base pairing, which may be advantageous in certain applications of the invention described herein.
[0258] In some embodiments of the invention, the RNA molecule may contain "nexus" and / or "hairpin" regions that may further specify the structure of the RNA molecule (see Briner et al., Molecular Cell (2014)).
[0259] As used herein, the term "direct repeat" refers to two or more repeats of a particular amino acid sequence of a nucleotide sequence.
[0260] As used herein, an RNA sequence or molecule that can "interact with" or "bind to" a CRISPR nuclease refers to the ability of the RNA sequence or molecule to form a CRISPR complex with a CRISPR nuclease.
[0261] As used herein, the term "operably linked" refers to a relationship (i.e., fusion, hybridization) between two sequences or molecules that allows them to function in their intended manner. In embodiments of the invention, when an RNA molecule is operably linked to a promoter, the RNA molecule and the promoter can function in their intended manner.
[0262] As used herein, the term "heterologous promoter" refers to a promoter that is not naturally present with the molecule being expressed or in the pathway being promoted.
[0263] As used herein, the "sequence identity" of a sequence or molecule is X% with respect to a second sequence or molecule if X% of the bases or amino acids between the sequences of the molecules are the same and in the same relative positions. For example, a first nucleotide sequence that has at least 95% sequence identity with a second nucleotide sequence has at least 95% base identity with the other sequence in the same relative positions.
[0264] nuclear localization sequence The terms "nuclear localization sequence" and "NLS" are used interchangeably to refer to an amino acid sequence / peptide that directs the transport of a bound protein from the cytoplasm across the nuclear membrane barrier. The term "NLS" is intended to encompass not only specific peptides that can direct the translocation of cytoplasmic polypeptides across the nuclear membrane barrier, but also their derivative nuclear localization sequences. An NLS can direct the nuclear translocation of a polypeptide by attaching it to the N-terminus, C-terminus, or both of the polypeptide. Furthermore, polypeptides with NLSs linked to the N- or C-terminus of an amino acid side chain randomly positioned in their amino acid sequence translocate. NLSs typically consist of one or more short sequences of positively charged lysines or arginines exposed on the protein surface, although other types of NLSs are known. Non-limiting examples of NLSs include NLS sequences derived from SV40 virus large T antigen, nucleoplasmin, c-myc, hRNPA1 M9 NLS, the IBB domain from importin alpha, fibroid T protein, human p53, mouse c-abl IV, influenza virus NS1, hepatitis virus delta antigen, mouse Mx1 protein, human poly(ADP-ribose) polymerase, and steroid hormone receptor (human) glucocorticoid.
[0265] delivery The CRISPR nucleases or CRISPR compositions described herein may be delivered as proteins, DNA molecules, RNA molecules, ribonucleoproteins (RNPs), nucleic acid vectors, or combinations thereof. In some embodiments, the RNA molecules comprise chemical modifications. Non-limiting examples of suitable chemical modifications include 2'-O-methyl (M), 2'-O-methyl-3'-phosphorothioate (MS) or 2'-O-methyl-3'-thioPACE (MSP), pseudouridine, and 1-methylpseudouridine. Each possibility represents a separate embodiment of this invention.
[0266] Nucleotide molecules such as CRISPR nucleases and / or polynucleotides encoding them, as described herein, and optionally additional proteins (e.g., ZFPs, TALENs, transcription factors, restriction enzymes) and / or guide RNAs, may be delivered to target cells by suitable means. Target cells may be any cell (e.g., eukaryotic, prokaryotic) in any environment (e.g., isolated or not, in culture, in vitro, ex vivo, in vivo, in planta).
[0267] In some embodiments, the composition to be delivered comprises a nuclease mRNA and a guide RNA. In some embodiments, the composition to be delivered comprises a nuclease mRNA, a guide RNA, and a donor template. In some embodiments, the composition to be delivered comprises a CRISPR nuclease and a guide RNA. In some embodiments, the composition to be delivered comprises a CRISPR nuclease, a guide RNA, and a donor template for gene editing, e.g., by homology-directed repair. In some embodiments, the composition to be delivered comprises a nuclease mRNA, a DNA-targeting RNA, and a tracrRNA. In some embodiments, the composition to be delivered comprises a nuclease mRNA, a DNA-targeting RNA, a tracrRNA, and a donor template. In some embodiments, the composition to be delivered comprises a CRISPR nuclease, a DNA-targeting RNA, and a tracrRNA. In some embodiments, the composition to be delivered comprises a CRISPR nuclease, a DNA-targeting RNA, a tracrRNA, and a donor template for gene editing, e.g., by homology-directed repair.
[0268] RNA compositions can be delivered using an appropriate viral vector system. Conventional viral and non-viral-based gene transfer methods can be used to introduce nucleic acids and / or CRISPR nucleases into cells (e.g., mammalian cells, plant cells, etc.) and target tissues. Such methods can also be used to provide encoded nucleic acids and / or CRISPR nuclease proteins to cells in vitro. In some embodiments, nucleic acids and / or CRISPR nucleases are administered for in vivo or ex vivo gene therapy. Non-viral vector delivery systems include naked nucleic acids and nucleic acids complexed with delivery vehicles such as liposomes or poloxamers. For reviews of gene therapy procedures, see Anderson, Science (1992); Nabel and Felgner, TIBTECH (1993); Mitani and Caskey, TIBTECH (1993); Dillon, TIBTECH (1993); Miller, Nature (1992); Van Brunt, Biotechnology (1988); Vigne et al., Restorative Neurology and Neuroscience 8:35-36 (1995); Kremer and Perricaudet, British Medical Bulletin (1995); Haddada et al., Current Topics in Microbiology and Immunology (1995) and Yu et al., Gene Therapy 1:13-26 (1994).
[0269] Non-viral methods for delivery of nucleic acids and / or proteins include electroporation, lipofection, microinjection, biolistics, particle gun acceleration, virosomes, liposomes, immunoliposomes, polycation or lipid:nucleic acid conjugates, artificial virions, and drug-enhanced nucleic acid uptake, or delivery into plant cells by bacteria or viruses (e.g., Agrobacterium, Rhizobium sp. NGR234, Sinorhizoboium meliloti, Mesorhizobium loti, Tobacco mosaic virus, Potato virus X, Cauliflower mosaic virus, Cassava vein mosaic virus). See, e.g., Chung et al. Trends Plant Sci. (2006). Sonoporation, e.g., using the Sonitron 2000 system (Rich-Mar), can also be used to deliver nucleic acids. Cationic lipid-mediated delivery of proteins and / or nucleic acids is also contemplated as an in vivo or in vitro delivery method (see Zuris et al., Nat. Biotechnol. (2015), Coelho et al., N. Engl. J. Med. (2013); Judge et al., Mol. Ther. (2006), and Basha et al., Mol. Ther. (2011).
[0270] Non-viral vectors, such as transposon-based systems (e.g., recombinant Sleeping Beauty transposon systems or recombinant PiggyBac transposon systems), may also be used to deliver and transpose the polynucleotide sequences of or encoding the molecules of the composition into target cells.
[0271] Other representative nucleic acid delivery systems include those offered by Amaxa® Biosystems (Cologne, Germany), Maxcyte, Inc. (Rockville, Md.), BTX Molecular Delivery Systems (Holliston, Mass.), and Copernicus Therapeutics Inc. (see, e.g., U.S. Pat. No. 6,008,336). Lipofectin is described, for example, in U.S. Pat. Nos. 5,049,386, 4,946,787, and 4,897,355, and lipofection reagents are commercially available (e.g., Transfectam®, Lipofectin®, and Lipofectamine® RNAiMAX). Cationic and neutral lipids suitable for efficient receptor-recognition lipofection of polynucleotides include those disclosed in WO 91 / 17424 and WO 91 / 16024. Delivery to cells (ex vivo) or target tissues (in vivo) is possible.
[0272] The preparation of lipid:nucleic acid complexes, including targeted liposomes such as immunolipid complexes, is widely known to those skilled in the art (see, e.g., Crystal, Science (1995); Blaese et al., Cancer Gene Ther. (1995); Behr et al., Bioconjugate Chem. (1994); Remy et al., Bioconjugate Chem. (1994); Gao and Huang, Gene Therapy (1995); Ahmad and Allen, Cancer Res., (1992); U.S. Patent Nos. 4,186,183; 4,217,344; 4,235,871; 4,261,975; 4,485,054; 4,501,728; 4,774,085; 4,837,028 and 4,946,787).
[0273] Another delivery method involves packaging the nucleic acid to be delivered in an EnGeneIC delivery vehicle (EDV). EDVs are specifically delivered to target tissues using bispecific antibodies, one arm of which has specificity for the target tissue and the other arm for the EDV. The antibody carries the EDV to the surface of the target cell, where it is then transported into the cell by endocytosis. Once inside the cell, its contents are released (see MacDiamid et al., Nature Biotechnology (2009)).
[0274] The use of RNA or DNA virus-based systems for nucleic acid delivery utilizes highly evolved methods to target viruses to specific cells in the body and transport the viral payload to the nucleus. Viral vectors can be administered directly to patients (in vivo) or used to treat cells in vitro, and the modified cells are then administered to patients (ex vivo). RNA or DNA virus-based systems for nucleic acid delivery include, but are not limited to, recombinant retroviruses, lentiviruses, adenoviruses, adeno-associated viruses, vaccinia viruses, and herpes simplex virus vectors for gene transfer. However, RNA viruses are preferred for delivery of the RNA compositions described herein. High transduction efficiencies have also been observed in various cells and target tissues. The nucleic acids of the present invention may be delivered by non-integrating lentiviruses. In some cases, lentivirus-mediated RNA delivery is utilized. In some cases, the lentivirus contains a nuclease mRNA and a guide RNA. In some cases, the lentivirus contains a nuclease mRNA, a guide RNA, and a donor template. In some cases, the lentivirus contains a nuclease protein and a guide RNA. In some cases, the lentivirus comprises a nuclease protein, a guide RNA, and / or a donor template for gene editing, e.g., by homology-directed repair. In some cases, the lentivirus comprises a nuclease mRNA, a DNA targeting RNA, and a tracrRNA. In some cases, the lentivirus comprises a nuclease mRNA, a DNA targeting RNA, a tracrRNA, and a donor template. In some cases, the lentivirus comprises a nuclease protein, a DNA targeting RNA, and a tracrRNA. In some cases, the lentivirus comprises a nuclease protein, a DNA targeting RNA, a tracrRNA, and a donor template for gene editing, e.g., by homology-directed repair.
[0275] As previously described, the compositions described herein can be delivered to target cells using non-integrating lentiviral particle methods (e.g., the LentiFlash® system). Such methods may also be used to deliver mRNA or other RNA to target cells, such that delivery of the RNA to the target cell results in assembly of the compositions described herein inside the target cell. See also WO 2013 / 014537, WO 2014 / 016690, WO 2016 / 185125, WO 2017 / 194902, and WO 2017 / 194903.
[0276] Retroviral tropism can be altered by incorporating foreign envelope proteins, expanding the potential target cell range. Lentiviral vectors are retroviral vectors that can transduce or infect non-dividing cells and typically produce high viral titers. The choice of retroviral gene transfer system depends on the target tissue. Retroviral vectors consist of cis-acting long terminal repeats that can package foreign sequences up to 6-10 kb. A minimal set of cis-acting LTRs is sufficient for vector replication and packaging, which is then used to integrate therapeutic genes into target cells and provide permanent transgene expression. Widely used retroviral vectors include those based on murine leukemia virus (MuLV), gibbon ape leukemia virus (GaLV), simian immunodeficiency virus (SIV), human immunodeficiency virus (HIV), and combinations thereof (see, e.g., Buchscher Panganiban, J. Virol. (1992); Johann et al., J. Virol. (1992); Sommerfelt et al., Virol. (1990); Wilson et al., J. Virol. (1989); Miller et al., J. Virol. (1991); WO 94 / 26877).
[0277] At least six viral vector approaches are currently available for gene transfer in clinical trials, using methods involving complementation of defective vectors by genes inserted into helper cell lines to generate transducing agents.
[0278] pLASN and MFG-S are examples of retroviral vectors used in clinical trials (Dunbar et al., Blood (1995); Kohn et al., Nat. Med. (1995); Malech et al., PNAS (1997)). PA317 / pLASN was the first therapeutic vector used in gene therapy (Blaese et al., Science (1995)). Transduction efficiencies of over 50% have been observed with MFG-S-packaged vectors (Ellem et al., Immunol Immunother. (1997); Dranoff et al., Hum. Gene Ther. (1997)).
[0279] Packaging cells are used to form viral particles capable of infecting host cells. Such cells include 293 cells, which package adenovirus (AAV), and psi.2 or PA317 cells, which package retrovirus. Viral vectors used in gene therapy are typically obtained by producer cell lines that package nucleic acid vectors into viral particles. The vector typically contains minimal viral sequences necessary for packaging and subsequent integration into the host (if applicable), with other viral sequences replaced by expression cassettes encoding the proteins to be expressed. Missing viral functions are supplied in trans by the packaging cell line. For example, AAV vectors used in gene therapy typically contain only the inverted terminal repeat (ITR) sequences of the AAV genome, which are necessary for packaging and integration into the host genome. The viral DNA is packaged into cell lines containing helper plasmids encoding other AAV genes, namely rep and cap, but lacking the ITR sequences. The cell lines are also infected with adenovirus as a helper. The helper virus facilitates AAV vector replication and expression of AAV genes from the helper plasmid. The helper plasmid lacks ITR sequences and is therefore not packaged in large quantities. Contamination with adenovirus can be reduced, for example, by heat treatment, to which adenovirus is more sensitive than AAV. Furthermore, AAV can be produced on a clinical scale using the baculovirus system (see U.S. Patent No. 7,479,554).
[0280] In many gene therapies, highly specific delivery of gene therapy vectors to specific tissues is desirable. Therefore, viral vectors can be engineered to have specificity for target cells by expressing a ligand as a fusion protein with the viral coat protein on the outer surface of the virus. The ligand is selected to have affinity for a receptor known to be present on the target cells. For example, Han et al., Proc. Natl. Acad. Sci. USA (1995) reported that Moloney murine leukemia virus can be engineered to express human heregulin fused to gp70 and that the recombinant virus infects specific human breast cancer cells expressing the human epidermal growth factor receptor. This principle can be extended to other virus-target cell combinations, where the target cells express a receptor and the virus expresses a fusion protein containing a ligand for the cell surface receptor. For example, filamentous phage can be engineered to display antibody fragments (e.g., FAB, Fv) that have substantially specific binding affinity for a selected cellular receptor. While this discussion primarily applies to viral vectors, the same principles can be applied to nonviral vectors. Such vectors can be modified to contain uptake sequences that facilitate uptake by specific target cells.
[0281] Gene therapy vectors can be delivered in vivo by administration to an individual patient, typically by systemic administration (e.g., intravenous, intraperitoneal, intramuscular, subcutaneous, or intracranial injection) or local application, as described below. Alternatively, vectors can be delivered ex vivo to cells, such as transplanted cells from an individual patient (e.g., lymphocytes, bone marrow aspirate, biopsy tissue) or hematopoietic stem cells from a universal donor, which are then re-implanted into the patient, typically after selection of cells that have incorporated the vector. In some embodiments, in vivo and ex vivo delivery of mRNA, as well as delivery of RNPs, may be utilized.
[0282] Ex vivo cell transfection for diagnostics, research, or gene therapy (e.g., by re-infusion of the transfected cells into the host organism) is widely known to those of skill in the art. In a preferred embodiment, cells are isolated from a subject organism, transfected with an RNA composition, and re-infused into the subject organism (e.g., a patient). A variety of cells suitable for ex vivo transfection are well known to those of skill in the art (see, e.g., Freshney, "Culture of Animal Cells, A Manual of Basic Technique and Specialized Applications" (6th edition, 2010) and the references cited therein for a discussion of methods for isolating and culturing cells from patients).
[0283] Suitable cells include, but are not limited to, eukaryotic and prokaryotic cells and / or cell lines. Non-limiting examples of such cells or cell lines generated from such cells include COS, CHO (e.g., CHO-S, CHO-K1, CHO-DG44, CHO-DUXB11, CHO-DUKX, CHOK1SV), VERO, MDCK, WI38, V79, B14AF28-G3, BHK, HaK, NSO, SP2 / 0-Ag14, HeLa, HEK293 (e.g., HEK293-F, HEK293-H, HEK293-T) and perC6 cells, plant cells (differentiated or undifferentiated), and insect cells such as Spodoptera fugitive flora (Sf), or fungal cells such as Saccharomyces, Pichia, and chizosaccharomyces. In some embodiments, the cell line is a CHO-K1, MDCK, or HEK293 cell line. Additionally, primary cells may be isolated, treated with a nuclease (e.g., ZFN or TALEN) or nuclease system (e.g., CRISPR), and then used ex vivo for reintroduction into a subject. Suitable primary cells include peripheral blood mononuclear cells (PBMCs) and blood cell subsets, such as, but not limited to, CD4+ T cells or CD8+ T cells. Suitable cells also include stem cells, such as, for example, embryonic stem cells, induced pluripotent stem cells, hematopoietic stem cells (CD34+), neural stem cells, and mesenchymal stem cells.
[0284] In one embodiment, stem cells are used in ex vivo procedures for cell transfection and gene therapy. The advantage of using stem cells is that they can be differentiated into other cells in vitro or introduced into a mammal (such as a cell donor) where they engraft in the bone marrow. Methods are known for differentiating CD34+ cells into clinically important immune cells in vitro using cytokines such as GM-CSF, IFNγ, and TNFα (see, for example, Inaba et al., J. Exp. Med. (1992)).
[0285] Stem cells are isolated for transduction and differentiation using known methods. For example, stem cells are isolated from bone marrow cells by panning the bone marrow cells with antibodies that bind to unwanted cells such as CD4+ and CD8+ (T cells), CD45+ (pan-B cells), GR-1 (granulocytes), and Iad (differentiated antigen-presenting cells) (see, for non-limiting examples, Inaba et al., J. Exp. Med. (1992)). In some embodiments, modified stem cells can also be used.
[0286] In particular, any one of the CRISPR nucleases described herein may be suitable for genome editing of post-mitotic cells or cells that are not actively dividing (e.g., arrested cells). Examples of post-mitotic cells that may be edited with the CRISPR nucleases of the invention include, but are not limited to, muscle cells, cardiomyocytes, hepatocytes, bone cells, and neurons.
[0287] Vectors (e.g., retroviruses, liposomes, etc.) containing therapeutic RNA compositions can also be administered directly to an organism for transduction of cells in vivo. Alternatively, naked RNA or mRNA can be administered. Administration can be by routes commonly used to introduce molecules with ultimate contact with blood or tissue cells, including, but not limited to, injection, infusion, topical application, and electroporation. Suitable methods for administering such nucleic acids are available and known to those of skill in the art, and while multiple routes of administration for a particular composition can be used, certain routes often result in more rapid and effective responses than others.
[0288] Suitable vectors for introducing transgenes into immune cells (e.g., T cells) include non-integrating lentiviral vectors, see, e.g., U.S. Patent Application Publication No. 2009 / 0117617.
[0289] Pharmaceutically acceptable carriers are determined in part by the composition being administered, as well as by the method used to administer the composition. Thus, there is a wide variety of suitable formulations of pharmaceutical compositions available, for example, as described in Remington's Pharmaceutical Sciences, 17th ed., 1989.
[0290] DNA repair by homologous recombination The term "homologous recombination repair" or "HDR" refers to a mechanism that repairs DNA damage in cells, e.g., during repair of double- and single-strand breaks in DNA. HDR requires nucleotide sequence homology and uses a "nucleic acid template" (the terms nucleic acid template and donor template are used interchangeably herein) to repair the sequence (e.g., DNA target sequence) where the double- or single-strand break occurred. This results in, for example, the transfer of genetic information from the nucleic acid template to the DNA target sequence. If the nucleic acid template sequence differs from the DNA target sequence and some or all of the nucleic acid template polynucleotide or oligonucleotide is incorporated into the DNA target sequence, HDR can result in an alteration (e.g., addition, deletion, mutation) of the DNA target sequence. In some embodiments, all or part of the nucleic acid template polynucleotide, or a copy of the nucleic acid template, is incorporated at the site of the DNA target sequence.
[0291] The terms "nucleic acid template" and "donor" refer to a nucleotide sequence to be inserted or copied into a genome. A nucleic acid template comprises, for example, one or more nucleotide sequences that may be added to a target nucleic acid, template a change in the target nucleic acid, or be used to modify a target sequence. The nucleic acid template sequence may be any length, for example, from 2 to 10,000 nucleotides (or any integer therebetween), preferably from about 100 to 1,000 nucleotides (or any integer therebetween), and more preferably from about 200 to 500 nucleotides. A nucleic acid template may be a single-stranded nucleic acid or a double-stranded nucleic acid. In some embodiments, a nucleic acid template comprises, for example, one or more nucleotide sequences corresponding to the wild-type sequence of a target nucleic acid, for example, at a target location. In some embodiments, a nucleic acid template comprises, for example, one or more ribonucleotide sequences corresponding to the wild-type sequence of a target nucleic acid, for example, at a target location. In some embodiments, a nucleic acid template comprises modified ribonucleotides.
[0292] Insertion of an exogenous sequence (also referred to as a "donor sequence," "donor template," or "donor") can also be performed, for example, to correct a mutant gene or increase expression of a wild-type gene. It is readily apparent that the donor sequence is usually not identical to the genomic sequence into which it is placed. The donor sequence can comprise a non-homologous sequence flanked by two homologous regions to enable efficient HDR at the target location. Furthermore, the donor sequence can comprise a vector molecule comprising a sequence that is not homologous to the target region in cellular chromatin. The donor molecule can comprise discontinuous regions homologous to cellular chromatin. For example, to target insertion of a sequence not normally present in the target region, the sequence can be present in the donor nucleic acid molecule and can be flanked by regions homologous to the sequence of the target region.
[0293] The donor polynucleotide may be single-stranded and / or double-stranded DNA or RNA and may be introduced into cells in a linear or circular form. See, for example, U.S. Patent Application Publication Nos. 2010 / 0047805; 2011 / 0281361; 2011 / 0207221; and 2019 / 0330620. When introduced in a linear form, the ends of the donor sequence can be protected (e.g., from exonuclease degradation) by methods known to those skilled in the art. For example, one or more dideoxynucleotide residues can be added to the 3' end of the linear molecule, and / or self-complementary oligonucleotides can be ligated to one or both ends. See, for example, Chang and Wilson, Proc. Natl. Acad. Sci. USA (1987); Nehls et al., Science (1996). Other methods of protecting exogenous polynucleotides from degradation include, but are not limited to, the addition of terminal amino groups and the use of modified internucleotide linkages (e.g., phosphorothioates, phosphoramidates, and O-methylribose or deoxyribose residues).
[0294] Thus, embodiments of the invention that use a donor template for repair may use single-stranded and / or double-stranded donor templates, DNA or RNA, that can be introduced into cells in linear or circular form. In some embodiments of the invention, the gene editing composition contains (1) an RNA molecule comprising a guide sequence that makes a double-stranded break in the gene prior to repair, and (2) a donor RNA template for repair, where the RNA molecule comprising the guide sequence is a first RNA molecule and the donor RNA template is a second RNA molecule. In some embodiments, the guide RNA molecule and the template RNA molecule are linked as part of a single molecule.
[0295] The donor sequence may be an oligonucleotide and may be used for gene correction or targeted modification of an endogenous sequence. The oligonucleotide may be introduced into a cell via a vector, electroporated into a cell, or by other methods known in the art. The oligonucleotide may be used to "correct" a mutant sequence in an endogenous gene (e.g., the sickle mutation of beta-globin) or may be used to insert a sequence for a desired purpose into an endogenous gene locus.
[0296] Polynucleotides can be introduced into cells as part of a vector molecule that contains additional sequences, such as, for example, an origin of replication, a promoter, and genes encoding antibiotic resistance. Additionally, donor polynucleotides can be introduced as naked nucleic acid, as nucleic acid complexed with agents such as liposomes, poloxamers, or delivered by recombinant viruses (e.g., adenovirus, AAV, herpesvirus, retrovirus, lentivirus, and integrase-deficient lentivirus (IDLV)).
[0297] The donor is generally inserted such that its expression is driven by the endogenous promoter of the integration site, i.e., the promoter that drives expression of the endogenous gene into which the donor is inserted. However, it will be apparent that the donor may also comprise a promoter and / or enhancer, e.g., a constitutive promoter or an inducible or tissue-specific promoter.
[0298] The donor molecule may be inserted into an endogenous gene such that all, a portion, or none of the endogenous gene is expressed. For example, a transgene described herein may be inserted into an endogenous locus such that a portion of the endogenous sequence (e.g., the N-terminus and / or C-terminus of the transgene) is expressed, e.g., as a fusion with the transgene, or none of the endogenous sequence is expressed. In other embodiments, a transgene (e.g., with or without additional coding sequence, e.g., an endogenous gene) is integrated into an endogenous locus, such as a safe harbor locus (e.g., the CCR5 gene, the CXCR4 gene, the PPP1R12c (also known as AAVS1) gene, the albumin gene, or the Rosa gene). See, e.g., U.S. Patent Nos. 7,951,925 and 8,110,379; U.S. Patent Application Publication Nos. 2008 / 0159996; 20100 / 0218264; 2010 / 0291048; 2012 / 0017290; 2011 / 0265198; 2013 / 0137104; 2013 / 0122591; 2013 / 0177983 and 2013 / 0177960, and U.S. Provisional Application No. 61 / 823,689).
[0299] When an endogenous sequence (either endogenous or a portion of a transgene) is expressed in conjunction with a transgene, the endogenous sequence may be a full-length sequence (wild-type or mutant) or a partial sequence. Preferably, the endogenous sequence is functional. Non-limiting examples of functions of these full-length or partial sequences include increasing the half-life of a polypeptide expressed by the transgene (e.g., a therapeutic gene) and / or acting as a carrier.
[0300] Additionally, although not essential for expression, the exogenous sequence may also include transcriptional or translational regulatory sequences, such as promoters, enhancers, insulators, internal ribosome entry sites, sequences encoding 2A peptides, and / or polyadenylation signals.
[0301] In some embodiments, the donor molecule comprises a sequence selected from the group consisting of a gene encoding a protein (e.g., a coding sequence encoding a protein that is missing in the cell or individual, or an alternative version of a gene encoding a protein), a regulatory sequence, and / or a sequence encoding a structural nucleic acid such as a microRNA or siRNA.
[0302] It is intended that the embodiments described above are applicable to one another, for example, it will be understood that an RNA molecule or composition of the invention may be utilized in a method of the invention.
[0303] All headings in this specification are for organizational purposes only and are not intended to limit the disclosure in any way. The content of each section is equally applicable to all sections.
[0304] Additional objects, advantages, and novel features of the present invention will become apparent to those skilled in the art upon examination of the following examples, which are not intended to be limiting. Additionally, each of the various embodiments and aspects of the invention as described hereinabove and as claimed below finds experimental support in the following examples.
[0305] It will be understood that features of the invention that are, for clarity, described in separate embodiments, may also be provided in combination in a single embodiment. Conversely, various features of the invention that are, for brevity, described in a single embodiment, may also be provided separately, in any suitable subcombination, or in other embodiments of the invention, as appropriate. Certain features described in various embodiments should not be considered essential features of those embodiments, unless the embodiment cannot function without those elements.
[0306] Generally, the nomenclature used herein and the laboratory procedures utilized in this invention include molecular, biochemical, microbiological, and recombinant DNA techniques. Such techniques are fully described in the literature. See, e.g., Sambrook et al., "Molecular Cloning: A Laboratory Manual" (1989); Ausubel, R.M. (Ed.), "Current Protocols in Molecular Biology" Volumes I-III (1994); Ausubel et al., "Current Protocols in Molecular Biology", John Wiley & Sons, Baltimore, Maryland (1989); Perbal, "A Practical Guide to Molecular Cloning", John Wiley & Sons, New York (1988); Watson et al., "Recombinant DNA", Scientific American Books, New York; Birren et al. (Eds.), "Genome Analysis: A Laboratory Manual Series", Vols. 1-4, Cold Spring Harbor Laboratory Press, New York. (1998); methods disclosed in U.S. Patent Nos. 4,666,828; 4,683,202; 4,801,531; 5,192,659 and 5,272,057; Cellis, JE (Ed.), "Cell Biology: A Laboratory Handbook", Volumes I-III (1994); Freshney, "Culture of Animal Cells - A Manual of Basic Technique" Third Edition, Wiley-Liss, NY (1994); Coligan JE (Ed.), "Current Protocols in Immunology" Volumes I-III (1994); Stites et al. (Eds.), "Basic and Clinical Immunology" (8th Edition), Appleton & Lange, Norwalk, CT (1994); Mishell and Shiigi (Eds.), "Strategies for Protein Purification and Characterization - A Laboratory Course Manual" CSHL Press (1996); Clokie and Kropinski (Eds.), "Bacteriophage Methods and Protocols", Volume 1: Isolation, Characterization, and Interactions (2009), all of which are incorporated by reference. Other general references are provided throughout this specification.
[0307] In order to facilitate a more complete understanding of the present invention, the following examples are provided. The following examples illustrate representative modes of making and practicing the invention. However, the scope of this invention is not limited to the specific embodiments disclosed in these examples, which are for illustrative purposes only. [Example]
[0308] Experiment details In order to facilitate a more complete understanding of the invention, the following examples are set forth. The following examples set forth exemplary modes of making and practicing the invention. However, the scope of the invention is not limited to the specific embodiments disclosed in these examples, which are for illustrative purposes only.
[0309] CRISPR repeats (crRNAs), transactivating RNAs (tracrRNAs) and nuclease polypeptides (OMNIs) were predicted from a metagenomic database of sequences from various environmental samples.
[0310] Construction of OMNI nuclease polypeptides To construct novel nuclease polypeptides (OMNIs), the open reading frames of several identified OMNIs were codon-optimized for expression in human cell lines, and the ORFs were cloned into the bacterial expression plasmid pET9a and the mammalian expression plasmid pmOMNI (Table 4).
[0311] sgRNA prediction and construction For each OMNI, we predicted the single guide RNA (sgRNA) by detecting CRISPR repeats and transactivating crRNAs in the respective bacterial genomes. In silico, we mapped the sequences of the native and immature crRNA and tracrRNA connected by the tetraloop "gaaa" and used RNA secondary structure prediction tools to predict the secondary structure elements of the duplex.
[0312] The predicted secondary structures of all duplex RNA elements (crRNA-tracrRNA chimeras) were used to identify potential tracrRNA sequences for sgRNA design. To overcome potential transcriptional and structural constraints and evaluate the flexibility of the sgRNA scaffolds in the human cellular environment, the nucleotide sequences of some candidate sgRNAs were slightly altered (Table 2, designated "v2" in the guide table). Finally, up to two versions of the candidate scaffolds designed for each OMNI were synthesized, attached downstream to a 22-nt universal and unique spacer sequence (T2, SEQ ID NO: 391), and cloned into a bacterial expression plasmid (pShuttleGuide, Table 4) under the control of an inducible T7 promoter combined with a U6 promoter for mammalian expression. T2 - GGAAGAGCAGAGCCTTGGTCTC (SEQ ID NO: 391)
[0313] In vitro depletion assay with TXTL In vitro PAM sequence depletion was tracked using the method described in Maxwell et al., Methods. 2018. Briefly, linear DNA expressing OMNI nuclease and a T7 promoter-driven sgRNA were added to an in vitro cell-free transcription / translation system (TXTL mix, Arbor Bioscience) along with a linear construct expressing T7 polymerase. RNA expression and protein translation in the TXTL mix resulted in the formation of ribonucleoprotein (RNP) complexes. Because linear DNA was used, a Chi6 DNA sequence was added to the TXTL reaction mix to inhibit the exonuclease activity of RecBCD and protect the linear DNA from degradation. The sgRNA spacer was designed to target a library of plasmids (pbPOS T2 library, Table 4) containing targeting protospacers flanking an 8N randomized set of potential PAM sequences. To add the necessary adapters and indexes to both the cleaved library and a control library expressing a non-targeting gRNA, depletion of PAM sequences in the libraries was measured using PCR followed by high-throughput sequencing. After deep sequencing, in vitro activity was confirmed by the percentage of depleted sequences with the same PAM sequence compared to their appearance in the control, indicating functional DNA cleavage by OMNI nuclease (Table 3).
[0314] Activity in human cells against endogenous genomic targets We also analyzed the ability of OMNIs to promote editing at specific genomic locations in human cells. To this end, for each OMNI, the corresponding OMNI-P2A-mCherry expression vector (pmOMNI, Table 4) was transfected into HeLa cells along with an sgRNA designed to target a specific location in the human genome (pShuttle Guide - Table 4, spacer sequence - Table 5). Cells were harvested 72 hours later. Half of the harvested cells were used to quantify transfection efficiency by FACS using mCherry fluorescence as a marker. The remaining half was lysed, and the genomic DNA was used to PCR amplify the corresponding putative genomic target. Next-generation sequencing (NGS) was performed on the amplicons, and the resulting sequences were used to calculate the percentage of editing at each target site. Short insertions or deletions (indels) surrounding the cut site are a typical result of DNA end repair after nuclease-induced DNA breaks. Therefore, the editing rate was estimated from the percentage of indels containing sequences within each amplicon.
[0315] The genomic activity of each OMNI was assessed using a panel of several unique sgRNAs designed to target different genomic locations. The results of these experiments are summarized in Table 5. As can be seen from the table (column 6, "Mean Maximum Activity"), all OMNIs show significant editing levels in human cells compared to the negative control (not shown).
[0316] [Table 1-1]
[0317] [Table 1-2]
[0318] [Table 1-3]
[0319] [Table 2-1]
[0320] Table 2-2
[0321] Table 2-3
[0322] Table 2-4
[0323] Table 2-5
[0324] Table 2-6
[0325] Table 2-7
[0326] Table 2-8
[0327] Table 2-9
[0328] Table 2-10
[0329] Table 2-11
[0330] Table 2-12
[0331] Table 2-13
[0332] Table 2-14
[0333] Table 2-15
[0334] Table 2-16
[0335] Table 2-17
[0336] Table 2-18
[0337] Table 2-19
[0338] Table 2-20
[0339] Table 2-21
[0340]
Table 2-22
[0341]
Table 3
[0342]
Table 4
[0343]
Table 5
[0344] References Ahmad and Allen (1992) “Antibody-mediated Specific Binging and Cytotoxicity of Lipsome-entrapped Doxorubicin to Lung Cancer Cells in Vitro”, Cancer Research 52:4817-20. Anderson (1992) “Human gene therapy”, Science 256:808-13. Basha et al. (2011) “Influence of Cationic Lipid Composition on Gene Silencing Properties of Lipid Nanoparticle Formulations of siRNA in Antigen-Presenting Cells”, Mol. Ther. 19(12):2186-200. Behr (1994) “Gene transfer with synthetic cationic amphiphiles: Prospects for gene therapy”, Bioconjuage Chem 5:382-89. Blaese et al. (1995) “Vectors in cancer therapy: how will they deliver”, Cancer Gene Ther. 2:291-97. Blaese et al. (1995) “T lympocyte-directed gene therapy for ADA-SCID: initial trial results after 4 years”, Science 270(5235):475-80. Briner et al. (2014) “Guide RNA functional modules direct Cas9 activity and orthognality”, Molecular Cell 56:333-39. Buchschacher and Panganiban (1992) “Human immunodeficiency virus vectors for inducible expression of foreign genes”, J. Virol. 66:2731-39. Burstein et al. (2017) “New CRISPR-Cas systems from uncultivated microbes”, Nature 542:237-41. Canver et al., (2015) "BCL11A enhancer dissection by Cas9-mediated in situ saturating mutagenesis", Nature Vol.527, Pgs.192-214. Chang and Wilson (1987) “Modification of DNA ends can decrease end-joining relative to homologous recombination in mammalian cells”, Proc. Natl. Acad. Sci. USA 84:4959- 4963. Charlesworth et al. (2019) “Identification of preexisting adaptive immunity to Cas9 proteins in humans”, Nature Medicine, 25(2), 249. Chung et al. (2006) “Agrobacterium is not alone: gene transfer to plants by viruses and other bacteria”, Trends Plant Sci. 11(1):1-4. Coelho et al. (2013) “Safety and efficacy of RNAi therapy for transthyretin amyloidosis” N. Engl. J. Med. 369, 819-829. Crystal (1995) “Transfer of genes to humans: early lessons and obstacles to success”, Science 270(5235):404-10. Dillon (1993) “Regulation gene expression in gene therapy” Trends in Biotechnology 11(5):167-173. Dranoff et al. (1997) “A phase I study of vaccination with autologous, irradiated melanoma cells engineered to secrete human granulocyte macrophage colony stimulating factor”, Hum. Gene Ther. 8(1):111-23. Dunbar et al. (1995) “Retrovirally marked CD34-enriched peripheral blood and bone marrow cells contribute to long-term engraftment after autologous transplantation”, Blood 85:3048-57. Ellem et al. (1997) “A case report: immune responses and clinical course of the first human use of ganulocyte / macrophage-colony-stimulating-factor-tranduced autologous melanoma cells for immunotherapy”, Cancer Immunol Immunother 44:10-20. Gao and Huang (1995) “Cationic liposome-mediated gene transfer” Gene Ther. 2(10):710-22. Haddada et al. (1995) “Gene Therapy Using Adenovirus Vectors”, in: The Molecular Repertoire of Adenoviruses III: Biology and Pathogenesis, ed. Doerfler and Boehm, pp. 297-306. Han et al. (1995) “Ligand-directed retro-viral targeting of human breast cancer cells”, Proc. Natl. Acad. Sci. USA 92(21):9747-51. Humbert et al., (2019) "Therapeutically relevant engraftment of a CRISPR-Cas9-edited HSC-enriched population with HbF reactivation in nonhuman primates", Sci. Trans. Med., Vol.11, Pgs.1-13. Inaba et al. (1992) “Generation of large numbers of dendritic cells from mouse bone marrow cultures supplemented with granulocyte / macrophage colony-stimulating factor”, J Exp Med. 176(6):1693-702. Jinek et al. (2012) “A programmable dual-RNA-guided DNA endonuclease in adaptive bacterial immunity”, Science 337(6096):816-21. Johan et al. (1992) “GLVR1, a receptor for gibbon ape leukemia virus, is homologous to a phosphate permease of Neurospora crassa and is expressed at high levels in the brain and thymus”, J Virol 66(3):1635-40. Judge et al. (2006) “Design of noninflammatory synthetic siRNA mediating potent gene silencing in vivo”, Mol Ther. 13(3):494-505. Kohn et al. (1995) “Engraftment of gene-modified umbilical cord blood cells in neonates with adnosine deaminase deficiency”, Nature Medicine 1:1017-23. Kremer and Perricaudet (1995) “Adenovirus and adeno-associated virus mediated gene transfer”, Br. Med. Bull. 51(1):31-44. Macdiarmid et al. (2009) “Sequential treatment of drug-resistant tumors with targeted minicells containing siRNA or a cytotoxic drug”, Nat Biotehcnol. 27(7):643-51. Malech et al. (1997) “Prolonged production of NADPH oxidase-corrected granulocyes after gene therapy of chronic granulomatous disease”, PNAS 94(22):12133-38. Maxwell et al. (2018) “A detailed cell-free transcription-translation-based assay to decipher CRISPR protospacer adjacent motifs”, Methods 14348-57 Miller et al. (1991) “Construction and properties of retrovirus packaging cells based on gibbon ape leukemia virus”, J Virol. 65(5):2220-24. Miller (1992) “Human gene therapy comes of age”, Nature 357:455-60. Mitani and Caskey (1993) “Delivering therapeutic genes - matching approach and application”, Trends in Biotechnology 11(5):162-66. Nabel and Felgner (1993) “Direct gene transfer for immunotherapy and immunization”, Trends in Biotechnology 11(5):211-15. Nehls et al. (1996) “Two genetically separable steps in the differentiation of thymic epithelium” Science 272:886-889. Remy et al. (1994) “Gene Transfer with a Series of Lipphilic DNA-Binding Molecules”, Bioconjugate Chem. 5(6):647-54. Sentmanat et al. (2018) “A Survey of Validation Strategies for CRISPR-Cas9 Editing”, Scientific Reports 8:888, doi:10.1038 / s41598-018-19441-8. Sommerfelt et al. (1990) “Localization of the receptor gene for type D simian retroviruses on human chromosome 19”, J. Virol. 64(12):6214-20. Van Brunt (1988) “Molecular framing: transgenic animals as bioactors” Biotechnology 6:1149-54. Vigne et al. (1995) “Third-generation adenovectors for gene therapy”, Restorative Neurology and Neuroscience 8(1,2): 35-36. Wagner et al. (2019) “High prevalence of Streptococcus pyogenes Cas9-reactive T cells within the adult human population” Nature Medicine, 25(2), 242 Wilson et al. (1989) “Formation of infectious hybrid virion with gibbon ape leukemia virus and human T-cell leukemia virus retroviral envelope glycoproteins and the gag and pol proteins of Moloney murine leukemia virus”, J. Virol. 63:2374-78. Yu et al. (1994) “Progress towards gene therapy for HIV infection”, Gene Ther. 1(1):13- 26. Zetsche et al. (2015) “Cpf1 is a single RNA-guided endonuclease of a class 2 CRIPSR-Cas system” Cell 163(3):759-71. Zuris et al. (2015) “Cationic lipid-mediated delivery of proteins enables efficient protein based genome editing in vitro and in vivo” Nat Biotechnol. 33(1):73-80.
Claims
1. A non-naturally occurring composition comprising a CRISPR nuclease comprising a sequence having at least 90% identity to a sequence selected from the group consisting of the amino acid sequences set forth in SEQ ID NOs: 12, 16, 18, 19, 1-11, 13-15, 17, and 20-22, or a nucleic acid molecule comprising a sequence encoding the CRISPR nuclease, wherein the CRISPR nuclease comprises one or more nuclear localization sequences (NLS).
2. 2. The composition of claim 1, further comprising one or more RNA molecules or a DNA polynucleotide encoding any one of said one or more RNA molecules, wherein said one or more RNA molecules and said CRISPR nuclease are not found together in nature, and said one or more RNA molecules are capable of forming a complex with said CRISPR nuclease and / or capable of directing said complex to a target site.
3. 3. The composition of claim 2, wherein the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO:1, and at least one RNA molecule comprises a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs:49-64.
4. 4. The composition of claim 3, wherein the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 1, and the at least one RNA molecule is a CRISPR RNA (crRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 50-53.
5. 5. The composition of claim 4, further comprising a trans-activating CRISPR RNA (tracrRNA) molecule comprising a sequence set forth in the group consisting of SEQ ID NOs: 54-64.
6. 3. The composition of claim 2, wherein the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO:1, and the at least one RNA molecule is a single guide RNA (sgRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs:49-64.
7. 3. The composition of claim 2, wherein the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 2, and at least one RNA molecule comprises a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 65-80.
8. 8. The composition of claim 7, wherein the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO:2, and the at least one RNA molecule is a CRISPR RNA (crRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs:66-69.
9. 9. The composition of claim 8, further comprising a trans-activating CRISPR RNA (tracrRNA) molecule comprising a sequence set forth in the group consisting of SEQ ID NOs: 70-80.
10. 3. The composition of claim 2, wherein the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO:2, and wherein at least one RNA molecule is a single guide RNA (sgRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs:65-80.
11. 3. The composition of claim 2, wherein the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO:3, and at least one RNA molecule comprises a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs:81-96.
12. 12. The composition of claim 11, wherein the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO:3, and the at least one RNA molecule is a CRISPR RNA (crRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs:82-85.
13. 13. The composition of claim 12, further comprising a trans-activating CRISPR RNA (tracrRNA) molecule comprising a sequence set forth in the group consisting of SEQ ID NOs: 86-96.
14. 3. The composition of claim 2, wherein the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO:3, and the at least one RNA molecule is a single guide RNA (sgRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs:81-96.
15. 3. The composition of claim 2, wherein the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO:4, and at least one RNA molecule comprises a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs:97-112.
16. 16. The composition of claim 15, wherein the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO:4, and the at least one RNA molecule is a CRISPR RNA (crRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs:98-101.
17. 17. The composition of claim 16, further comprising a trans-activating CRISPR RNA (tracrRNA) molecule comprising a sequence set forth in the group consisting of SEQ ID NOs: 102-112.
18. 3. The composition of claim 2, wherein the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO:4, and the at least one RNA molecule is a single guide RNA (sgRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs:97-112.
19. 3. The composition of claim 2, wherein the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO:5, and at least one RNA molecule comprises a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs:113-128.
20. 20. The composition of claim 19, wherein the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO:5, and the at least one RNA molecule is a CRISPR RNA (crRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs:114-117.
21. 21. The composition of claim 20, further comprising a trans-activating CRISPR RNA (tracrRNA) molecule comprising a sequence set forth in the group consisting of SEQ ID NOs: 118-128.
22. 3. The composition of claim 2, wherein the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO:5, and the at least one RNA molecule is a single guide RNA (sgRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs:113-128.
23. 3. The composition of claim 2, wherein the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO:6, and at least one RNA molecule comprises a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs:129-144.
24. 24. The composition of claim 23, wherein the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO:6, and the at least one RNA molecule is a CRISPR RNA (crRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs:130-133.
25. 25. The composition of claim 24, further comprising a trans-activating CRISPR RNA (tracrRNA) molecule comprising a sequence set forth in the group consisting of SEQ ID NOs: 134-144.
26. 3. The composition of claim 2, wherein the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO:6, and the at least one RNA molecule is a single guide RNA (sgRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs:129-144.
27. 3. The composition of claim 2, wherein the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 7, and at least one RNA molecule comprises a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 145-163.
28. 28. The composition of claim 27, wherein the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 7, and the at least one RNA molecule is a CRISPR RNA (crRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 146-149.
29. 29. The composition of claim 28, further comprising a trans-activating CRISPR RNA (tracrRNA) molecule comprising a sequence set forth in the group consisting of SEQ ID NOs: 150-160, 162 and 163.
30. 3. The composition of claim 2, wherein the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 7, and the at least one RNA molecule is a single guide RNA (sgRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 145-163.
31. 3. The composition of claim 2, wherein the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO:8, and at least one RNA molecule comprises a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs:164-179.
32. 32. The composition of claim 31, wherein the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO:8, and the at least one RNA molecule is a CRISPR RNA (crRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs:165-168.
33. 33. The composition of claim 32, further comprising a trans-activating CRISPR RNA (tracrRNA) molecule comprising a sequence set forth in the group consisting of SEQ ID NOs: 169-179.
34. 3. The composition of claim 2, wherein the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO:8, and the at least one RNA molecule is a single guide RNA (sgRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs:164-179.
35. 3. The composition of claim 2, wherein the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO:9, and at least one RNA molecule comprises a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs:180-195.
36. 36. The composition of claim 35, wherein the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO:9, and the at least one RNA molecule is a CRISPR RNA (crRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs:181-184.
37. 37. The composition of claim 36, further comprising a trans-activating CRISPR RNA (tracrRNA) molecule comprising a sequence set forth in the group consisting of SEQ ID NOs: 185-195.
38. The composition of claim 2, wherein the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 9, and the at least one RNA molecule is a single guide RNA (sgRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 180 to 195.
39. 3. The composition of claim 2, wherein the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 10, and at least one RNA molecule comprises a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 196-211.
40. 40. The composition of claim 39, wherein the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 10, and the at least one RNA molecule is a CRISPR RNA (crRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 197-200.
41. 41. The composition of claim 40, further comprising a trans-activating CRISPR RNA (tracrRNA) molecule comprising a sequence set forth in the group consisting of SEQ ID NOs: 201-211.
42. 3. The composition of claim 2, wherein the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 10, and the at least one RNA molecule is a single guide RNA (sgRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 196-211.
43. 3. The composition of claim 2, wherein the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 11, and at least one RNA molecule comprises a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 212-230.
44. 44. The composition of claim 43, wherein the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 11, and the at least one RNA molecule is a CRISPR RNA (crRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 213-216.
45. 45. The composition of claim 44, further comprising a trans-activating CRISPR RNA (tracrRNA) molecule comprising a sequence set forth in the group consisting of SEQ ID NOs: 217-225 and 227-230.
46. 3. The composition of claim 2, wherein the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 11, and the at least one RNA molecule is a single guide RNA (sgRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 212-230.
47. 3. The composition of claim 2, wherein the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 12, and at least one RNA molecule comprises a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 231-249.
48. 48. The composition of claim 47, wherein the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 12, and the at least one RNA molecule is a CRISPR RNA (crRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 232-235.
49. 49. The composition of claim 48, further comprising a trans-activating CRISPR RNA (tracrRNA) molecule comprising a sequence set forth in the group consisting of SEQ ID NOs: 236-244 and 246-249.
50. 3. The composition of claim 2, wherein the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 12, and the at least one RNA molecule is a single guide RNA (sgRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 231-249.
51. 3. The composition of claim 2, wherein the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 13, and at least one RNA molecule comprises a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 250-270.
52. 52. The composition of claim 51, wherein the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 13, and wherein the at least one RNA molecule is a CRISPR RNA (crRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 251-254.
53. 53. The composition of claim 52, further comprising a trans-activating CRISPR RNA (tracrRNA) molecule comprising a sequence set forth in the group consisting of SEQ ID NOs: 255-265 and 267-270.
54. 3. The composition of claim 2, wherein the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 13, and the at least one RNA molecule is a single guide RNA (sgRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 250-270.
55. 3. The composition of claim 2, wherein the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 14, and at least one RNA molecule comprises a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 271-286.
56. 56. The composition of claim 55, wherein the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 14, and the at least one RNA molecule is a CRISPR RNA (crRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 272-275.
57. 57. The composition of claim 56, further comprising a trans-activating CRISPR RNA (tracrRNA) molecule comprising a sequence set forth in the group consisting of SEQ ID NOs: 276-286.
58. 3. The composition of claim 2, wherein the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 14, and the at least one RNA molecule is a single guide RNA (sgRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 271-286.
59. 3. The composition of claim 2, wherein the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 15, and at least one RNA molecule comprises a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 287-302.
60. 60. The composition of claim 59, wherein the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 15, and the at least one RNA molecule is a CRISPR RNA (crRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 288-291.
61. 61. The composition of claim 60, further comprising a trans-activating CRISPR RNA (tracrRNA) molecule comprising a sequence set forth in the group consisting of SEQ ID NOs: 292-302.
62. 3. The composition of claim 2, wherein the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 15, and the at least one RNA molecule is a single guide RNA (sgRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 287-302.
63. 3. The composition of claim 2, wherein the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 16, and at least one RNA molecule comprises a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 303-318.
64. 64. The composition of claim 63, wherein the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 16, and wherein the at least one RNA molecule is a CRISPR RNA (crRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 304-307.
65. 65. The composition of claim 64, further comprising a trans-activating CRISPR RNA (tracrRNA) molecule comprising a sequence set forth in the group consisting of SEQ ID NOs: 308-318.
66. 3. The composition of claim 2, wherein the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 16, and the at least one RNA molecule is a single guide RNA (sgRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 303-318.
67. 3. The composition of claim 2, wherein the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 17, and at least one RNA molecule comprises a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 319-332.
68. 68. The composition of claim 67, wherein the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 17, and the at least one RNA molecule is a CRISPR RNA (crRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 320-323.
69. 69. The composition of claim 68, further comprising a trans-activating CRISPR RNA (tracrRNA) molecule comprising a sequence set forth in the group consisting of SEQ ID NOs: 324-332.
70. 3. The composition of claim 2, wherein the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 17, and the at least one RNA molecule is a single guide RNA (sgRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 319-332.
71. 3. The composition of claim 2, wherein the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 18, and at least one RNA molecule comprises a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 333-353.
72. 72. The composition of claim 71, wherein the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 18, and wherein the at least one RNA molecule is a CRISPR RNA (crRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 334-337.
73. 73. The composition of claim 72, further comprising a trans-activating CRISPR RNA (tracrRNA) molecule comprising a sequence set forth in the group consisting of SEQ ID NOs: 338-348 and 350-353.
74. 3. The composition of claim 2, wherein the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 18, and the at least one RNA molecule is a single guide RNA (sgRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 333-353.
75. 3. The composition of claim 2, wherein the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 19, and at least one RNA molecule comprises a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 354-369.
76. 76. The composition of claim 75, wherein the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 19, and wherein the at least one RNA molecule is a CRISPR RNA (crRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 355-358.
77. 77. The composition of claim 76, further comprising a trans-activating CRISPR RNA (tracrRNA) molecule comprising a sequence set forth in the group consisting of SEQ ID NOs: 359-369.
78. 3. The composition of claim 2, wherein the CRISPR nuclease comprises a sequence having at least 90% identity to the amino acid sequence set forth in SEQ ID NO: 19, and the at least one RNA molecule is a single guide RNA (sgRNA) molecule comprising a guide sequence portion and a sequence selected from the group consisting of the sequences set forth in SEQ ID NOs: 354-369.
79. the amino acid sequence of the CRISPR nuclease comprises a conserved insertion consisting of 343 consensus sites, each of which may be an amino acid or gap such that removing the gap and joining amino acids results in the entire conserved insertion sequence; Consensus site 1 of the conserved insertion region is an asparagine (N) residue; Consensus site 4 of the conserved insertion region is a leucine (L) residue; The consensus site 45 of the conserved insertion region is a lysine (K) residue; The consensus site 50 of the conserved insertion region is a proline (P) residue; The consensus site 61 of the conserved insertion region is a valine (V) residue; The consensus site 63 of the conserved insertion region is a valine (V) residue; The consensus site 87 of the conserved insertion region is a leucine (L) residue; The consensus site 94 of the conserved insertion region is a glycine (G) residue; The consensus site 162 of the conserved insertion region is a glycine (G) residue; The consensus site 164 of the conserved insertion region is a tyrosine (Y) residue; The consensus site 165 of the conserved insertion region is a tryptophan (W) residue; The consensus site 184 of the conserved insertion region is an arginine (R) residue; The consensus site 197 of the conserved insertion region is a glycine (G) residue; The consensus site 204 of the conserved insertion region is a phenylalanine (F) residue; The consensus site 232 of the conserved insertion region is a leucine (L) residue; The consensus site 241 of the conserved insertion region is a phenylalanine (F) residue; The consensus site 243 of the conserved insertion region is a lysine (K) residue; The consensus site 249 of the conserved insertion region is a proline (P) residue; The consensus site 258 of the conserved insertion region is a leucine (L) residue; The consensus site 273 of the conserved insertion region is an aspartic acid (D) residue; The consensus site 280 of the conserved insertion region is a glutamine (Q) residue; The consensus site 291 of the conserved insertion region is a tyrosine (Y) residue; The consensus site 316 of the conserved insertion region is a proline (P) residue; the consensus site 327 of the conserved insertion region is a glycine (G) residue; and / or The consensus site 343 of the conserved insertion region is a proline (P) residue.
79. The composition of any one of claims 1 to 78.
80. the amino acid sequence of the CRISPR nuclease comprises a conserved insertion consisting of 343 consensus sites, each of which may be an amino acid or gap such that removing the gap and joining amino acids results in the entire conserved insertion sequence; Consensus site 1 of the conserved insertion region is an asparagine (N) residue; Consensus site 2 of the conserved insertion region is a lysine (K), asparagine (N), or serine (S) residue; Consensus site 3 of the conserved insertion region is an isoleucine (I), methionine (M), or valine (V) residue; Consensus site 4 of the conserved insertion region is an isoleucine (I), leucine (L), or valine (V) residue; Consensus site 8 of the conserved insertion region is a proline (P), serine (S), or tyrosine (Y) residue; The consensus site 19 of the conserved insertion region is a cysteine (C), isoleucine (I), or leucine (L) residue, or a gap; The consensus site 29 of the conserved insertion region is an isoleucine (I), valine (V), tryptophan (W), or tyrosine (Y) residue; The consensus site 50 of the conserved insertion region is a proline (P) residue; The consensus site 56 of the conserved insertion region is an aspartic acid (D), glutamic acid (E), or glutamine (Q) residue; The consensus site 65 of the conserved insertion region is a glutamic acid (E), glycine (G), or asparagine (N) residue, or a gap; The consensus site 66 of the conserved insertion region is a glycine (G), histidine (H), lysine (K), asparagine (N), or threonine (T) residue; The consensus site 76 of the conserved insertion region is a glutamic acid (E) or arginine (R) residue, or a gap; The consensus site 77 of the conserved insertion region is a histidine (H) or leucine (L) residue, or a gap; The consensus site 87 of the conserved insertion region is a glutamic acid (E), leucine (L), or arginine (R) residue; The consensus site 89 of the conserved insertion region is a glycine (G), leucine (L), or glutamine (Q) residue, or a gap; The consensus site 90 of the conserved insertion region is an aspartic acid (D), lysine (K), asparagine (N), or serine (S) residue, or a gap; The consensus site 91 of the conserved insertion region is a glycine (G), asparagine (N), or tyrosine (Y) residue, or a gap; The consensus site 92 of the conserved insertion region is a glutamic acid (E), phenylalanine (F), or valine (V) residue, or a gap; The consensus site 94 of the conserved insertion region is a glycine (G) residue; The consensus site 96 of the conserved insertion region is an alanine (A), histidine (H), or tyrosine (Y) residue; The consensus site 100 of the conserved insertion region is a leucine (L), valine (V), or tyrosine (Y) residue; The consensus site 105 of the conserved insertion region is a glycine (G), lysine (K), or asparagine (N) residue, or a gap; The consensus site 113 of the conserved insertion region is a phenylalanine (F), histidine (H), lysine (K), or tyrosine (Y) residue; The consensus site 117 of the conserved insertion region is an alanine (A) or proline (P) residue; The consensus site 125 of the conserved insertion region is an isoleucine (I), arginine (R), or valine (V) residue; The consensus site 128 of the conserved insertion region is an alanine (A), glycine (G), threonine (T), or valine (V) residue; The consensus site 132 of the conserved insertion region is a leucine (L) residue or a gap; The consensus site 133 of the conserved insertion region is an asparagine (N) residue or a gap; The consensus site 134 of the conserved insertion region is a lysine (K) or arginine (R) residue, or a gap; The consensus site 135 of the conserved insertion region is an aspartic acid (D) residue or a gap; The consensus site 136 of the conserved insertion region is a glycine (G) or serine (S) residue, or a gap; The consensus site 140 of the conserved insertion region is a phenylalanine (F) or leucine (L) residue; The consensus site 147 of the conserved insertion region is a phenylalanine (F) residue or a gap; The consensus site 156 of the conserved insertion region is an aspartic acid (D), arginine (R), or serine (S) residue, or a gap; The consensus site 162 of the conserved insertion region is a glycine (G) or serine (S) residue; The consensus site 164 of the conserved insertion region is a cysteine (C) or tyrosine (Y) residue; The consensus site 165 of the conserved insertion region is a lysine (K) or tryptophan (W) residue; The consensus site 176 of the conserved insertion region is a valine (V) or tryptophan (W) residue, or a gap; The consensus site 186 of the conserved insertion region is an alanine (A), lysine (K), or arginine (R) residue, or a gap; The consensus site 187 of the conserved insertion region is a threonine (T) residue or a gap; The consensus site 195 of the conserved insertion region is a leucine (L) or valine (V) residue; The consensus site 197 of the conserved insertion region is a glycine (G) residue; The consensus site 199 of the conserved insertion region is an isoleucine (I), leucine (L), threonine (T), or valine (V) residue; The consensus site 204 of the conserved insertion region is a phenylalanine (F) residue; The consensus site 208 of the conserved insertion region is a proline (P) residue or a gap; The consensus site 209 of the conserved insertion region is an aspartic acid (D) residue or a gap; The consensus site 211 of the conserved insertion region is a phenylalanine (F) or tyrosine (Y) residue; The consensus site 215 of the conserved insertion region is a glutamic acid (E) or valine (V) residue, or a gap; The consensus site 216 of the conserved insertion region is an aspartic acid (D) or glutamic acid (E) residue, or a gap; The consensus site 217 of the conserved insertion region is an asparagine (N) or serine (S) residue, or a gap; The consensus site 218 of the conserved insertion region is a serine (S) residue or a gap; The consensus site 219 of the conserved insertion region is a valine (V) residue or a gap; The consensus site 225 of the conserved insertion region is an aspartic acid (D), histidine (H), or proline (P) residue; The consensus site 226 of the conserved insertion region is a glycine (G) or arginine (R) residue; The consensus site 235 of the conserved insertion region is an aspartic acid (D), glutamic acid (E), phenylalanine (F), methionine (M), or asparagine (N) residue; The consensus site 241 of the conserved insertion region is a phenylalanine (F), isoleucine (I), or tyrosine (Y) residue; The consensus site 249 of the conserved insertion region is a glycine (G), proline (P), or serine (S) residue; The consensus site 260 of the conserved insertion region is an alanine (A) or glycine (G) residue; The consensus site 268 of the conserved insertion region is an alanine (A), phenylalanine (F), or leucine (L) residue; The consensus site 273 of the conserved insertion region is an aspartic acid (D) or asparagine (N) residue; The consensus site 274 of the conserved insertion region is a glycine (G) residue or a gap; The consensus site 288 of the conserved insertion region is a glycine (G), serine (S), or threonine (T) residue; The consensus site 289 of the conserved insertion region is a lysine (K), glutamine (Q), or arginine (R) residue; The consensus site 290 of the conserved insertion region is a histidine (H) or tyrosine (Y) residue; The consensus site 303 of the conserved insertion region is a glutamic acid (E) or lysine (K) residue, or a gap; The consensus site 304 of the conserved insertion region is a glutamic acid (E) residue or a gap; The consensus site 305 of the conserved insertion region is a glutamine (Q) or arginine (R) residue, or a gap; The consensus site 306 of the conserved insertion region is a valine (V) residue or a gap; The consensus site 307 of the conserved insertion region is a threonine (T) residue or a gap; The consensus site 316 of the conserved insertion region is a proline (P) or glutamine (Q) residue; The consensus site 326 of the conserved insertion region is an aspartic acid (D) or glutamic acid (E) residue; The consensus site 327 of the conserved insertion region is a glycine (G) residue; The consensus site 331 of the conserved insertion region is an alanine (A) or leucine (L) residue, or a gap; The consensus site 333 of the conserved insertion region is an aspartic acid (D) or asparagine (N) residue; The consensus site 338 of the conserved insertion region is a glutamic acid (E), leucine (L), or arginine (R) residue; The consensus site 339 of the conserved insertion region is an alanine (A), isoleucine (I), or tyrosine (Y) residue; The consensus site 341 of the conserved insertion region is a phenylalanine (F) or leucine (L) residue; and / or The consensus site 343 of the conserved insertion region is a proline (P) or serine (S) residue.
79. The composition of any one of claims 1 to 78.
81. The composition of any one of claims 1 to 80, wherein the amino acid sequence of the CRISPR nuclease is other than SEQ ID NOs: 45 to 48.
82. 82. The composition of any one of claims 1 to 81, wherein the CRISPR nuclease is a nickase having an inactivated RuvC domain created by amino acid substitutions in the CRISPR nuclease at the positions shown in column 5 of Table 1.
83. 82. The composition of any one of claims 1 to 81, wherein the CRISPR nuclease is a nickase having an inactivating HNH domain created by amino acid substitutions in the CRISPR nuclease at the positions shown in column 6 of Table 1.
84. 82. The composition of any one of claims 1 to 81, wherein the CRISPR nuclease is an inactivated nuclease having an inactivated RuvC domain and an inactivated HNH domain created by substitution of the CRISPR nuclease at the positions shown in column 7 of Table 1.
85. 24. A method for modifying the nucleotide sequence of a DNA target site in the genome of a cell-free system or a cell, the method comprising introducing into said cell a composition according to any one of claims 2 to 23.
86. 86. The method of claim 85, wherein the CRISPR nuclease cleaves the DNA strand adjacent to a protospacer adjacent motif (PAM) sequence of a CRISPR nuclease shown in columns 2-3 of Table 3, and / or cleaves the DNA strand adjacent to a sequence complementary to the PAM sequence.
87. 86. The method of Claim 85, wherein the CRISPR nuclease is a nickase having an inactivated RuvC domain created by amino acid substitutions in the CRISPR nuclease at the positions shown in column 5 of Table 1, and cleaves the DNA strand adjacent to the sequence complementary to the PAM sequence.
88. 86. The method of Claim 85, wherein the CRISPR nuclease is a nickase having an inactivating HNH domain created by amino acid substitutions in the CRISPR nuclease at the positions shown in column 6 of Table 1, and cleaves the DNA strand adjacent to the PAM sequence.
89. 89. The method of any one of claims 85 to 88, wherein the cell is a eukaryotic cell or a prokaryotic cell.
90. 90. The method of claim 89, wherein the cell is a mammalian cell.
91. 91. The method of claim 90, wherein the cell is a human cell.