Artificial expression constructs for regulating gene expression in dopaminergic neurons
Patent Information
- Application Number
- JP2024550135
- Authority / Receiving Office
- JP · JP
- Patent Type
- Applications
- Current Assignee / Owner
- Priority Date
- 2022-02-24
- Filing Date
- 2023-02-24
- Publication Date
- 2026-02-10
AI Technical Summary
Current methods for labeling and manipulating specific types of central nervous system cells, particularly dopaminergic neurons, are costly and inefficient, requiring complex animal models and being unsuitable for human application.
Development of artificial expression constructs that utilize specific enhancer elements, such as eHGT_888m and eHGT_606h, to induce gene expression in dopaminergic neurons, allowing for targeted modulation of gene expression.
The artificial expression constructs enable rapid and high-level expression of transgenes in dopaminergic neurons, overcoming the limitations of existing technologies and providing a more efficient and specific tool for research and potential therapeutic applications.
Smart Images

Figure 00000000_0000_ABST
Abstract
Description
[Technical field]
[0001] CROSS-REFERENCE TO RELATED APPLICATIONS This application claims priority to U.S. Provisional Patent Application No. 63 / 313,653, filed February 24, 2022, the contents of which are incorporated by reference in their entirety as if set forth herein.
[0002] The present disclosure provides artificial expression constructs for regulating gene expression in targeted types of central nervous system cells. The artificial expression constructs of the present invention can be used to express synthetic genes or regulate gene expression in dopaminergic neurons of the central nervous system.
[0003] Sequence Listing Reference The sequence listing accompanying this application is provided in XML format rather than hard copy, and is incorporated herein by reference. The XML file containing this sequence listing is named A166-0035PCT.xml. This file is 152KB in size, was created on February 23, 2023, and was submitted electronically via the Patent Center. [Background technology]
[0004] To fully understand the biology of the brain, it is necessary to distinguish between different cell types, define them, and investigate them in detail, as well as to identify artificial expression constructs that can label and perturb those cells. In mice, driver lines expressing recombinases have been used successfully to label cell populations that share marker gene expression. However, the creation, maintenance, and use of such lines that can label specific cell types with high specificity is costly and often requires crossing transgenic animals between three species, which only results in low frequency of obtaining the desired experimental animals. Moreover, these tools cannot be applied to humans because of the need for germline transgenic animals. Summary of the Invention [Means for solving the problem]
[0005] The present disclosure provides artificial expression constructs that direct gene expression in dopaminergic neurons.
[0006] In certain embodiments of the artificial expression constructs of the present disclosure, an enhancer selected from eHGT_888m, eHGT_897m, eHGT_606h, MGT_E68, eHGT_889h, eHGT_1038m, eHGT_589m, eHGT_647m and eHGT_483m is utilized to induce gene expression in dopaminergic neurons.
[0007] In certain embodiments, these artificial enhancer elements are concatemerized enhancer cores.Examples of such enhancer cores include eHGT_606h, eHGT_1039h, hTH, eHGT_888m, eHGT_888h, eHGT_017m and / or MGT_E51 cores, or concatemerized these cores.By using such artificial enhancer elements, transgenes can be rapidly expressed and high expression can be obtained, compared to using the full length of original (natural) enhancer alone.
[0008] In certain embodiments, the enhancer cores include sequences shown in SEQ ID NO:2, SEQ ID NO:4, SEQ ID NO:6, SEQ ID NO:8, SEQ ID NO:83, SEQ ID NO:85, SEQ ID NO:90, or SEQ ID NO:88. In certain embodiments, these enhancer cores are concatemerized and have 2, 3, 4, 5, 6, 7, 8, 9, or 10 copies of the core sequence. In certain embodiments, concatemers containing 3 copies of the selected enhancer core include sequences shown in any of SEQ ID NO:3, SEQ ID NO:5, SEQ ID NO:7, SEQ ID NO:9, SEQ ID NO:84, SEQ ID NO:86, SEQ ID NO:91, and SEQ ID NO:89.
[0009] Certain embodiments of the artificial expression constructs of the present disclosure utilize 3xCore2-MGT_E51, 3xcore4_eHGT_606h, 3xcore5_eHGT_606h, 3xcore_eHGT_1039h, 3xCore_hTH, core4_eHGT_888m, 3Xcore4_eHGT_888m, 3Xcore4_eHGT_888h and / or 3xcore1_eHGT_017m to selectively induce gene expression in dopaminergic neurons.
[0010] In certain embodiments, artificial expression constructs are provided that include features of vectors described herein, such as vectors selected from CN3406, CN3509, AiP1417, CN2436, AiP1240, CN3739, CN3889, CN2839, CN2847, CN2431, CN3058, CN3059, CN3829, CN3283, CN3863, CN4365, CN4830, and CN3551. [Brief description of the drawings]
[0011] Some of the drawings submitted in this application may be more easily understood in color, and applicants hereby contemplate color versions of these drawings as part of the original application and reserve the right to submit color images of such drawings in subsequent proceedings.
[0012] [Figure 1] An inverted epifluorescence microscope image showing native SYFP2 expression in the sagittal plane of mouse brain 28 days after retro-orbital delivery of 6.0×1011 viral genome copies of AAV vector #CN3406 (eHGT_888m). Scale bar: 1 mm.
[0013] [Figure 2A-2B](Figure 2A) Epifluorescence microscopy image showing co-staining of anti-green fluorescent protein (GFP) (for detection of viral expression of SYFP2) and anti-tyrosine hydroxylase (TH) (as a marker for dopaminergic neurons) in the sagittal plane of mouse brain 28 days after retro-orbital delivery of AAV vector #CN3406 (eHGT_888m) with 6.0x1011 viral genome copies. Scale bar: 1 mm. (Figure 2B) Enlarged image of the boxed area shown in Figure 2A, showing co-localization of GFP and TH signals in midbrain dopamine neurons. Scale bar: 200 microns.
[0014] [Figure 3A-3E] Conserved enhancer activity in midbrain dopamine neurons of adult mice after in vivo retro-orbital injection of CN3863 of PHP.eB serotype at a dose of 5x1011 viral genome copies. (Fig. 3A) Midsagittal mouse brain section showing enrichment of native SYFP2 expression in the substantia nigra pars compacta (SNc). (Fig. 3B) Diagram comparing the 1,293 bp eHGT_888m full-length genomic fragment (CN3406) with the smaller 287 bp core4-eHGT_888m fragment (CN3863), both of which confer strong fluorescent reporter expression in midbrain dopamine neurons. Other partially overlapping fragments, core1, core2, and core3, showed no dopamine neuron activity when tested in vivo. (Figure 3C) Anti-tyrosine hydroxylase (anti-TH) immunostaining of a coronal brain section showing dopaminergic neurons in the midbrain. (Figure 3D) Image of the same field showing anti-GFP immunostaining of neurons labeled with the Enhancer virus. (Figure 3E) Overlay image showing strong co-labeling of the SNc region, strongly suggesting that the virally labeled neurons are dopaminergic neurons. Abbreviations: VTA: ventral tegmental area; SNr: substantia nigra pars reticulata; SNc: substantia nigra pars compacta. Scale bars: (Figure 3A) 1 mm, (Figure 3C-3E) 300 microns.
[0015] [Figure 4A-4C] Figure 4 shows preserved enhancer activity in midbrain dopamine neurons of adult rhesus monkeys after in vivo injection of CN3863, a serotype of PHP.eB, targeted to the substantia nigra pars compacta (SNc) region. (Figure 4A) Anti-tyrosine hydroxylase (anti-TH) immunostaining of a coronal brain section showing midbrain dopaminergic neurons. (Figure 4B) Image of the same field showing anti-GFP immunostaining of neurons labeled with the enhancer virus. (Figure 4C) Overlay image showing strong co-labeling of the SNc region, strongly suggesting that the virus-labeled neurons are dopaminergic neurons. Asterisks indicate rare virus-labeled neurons that were negative for TH. Dotted lines indicate the border of the SNc region. Abbreviations: VTA: ventral tegmental area; SNr: substantia nigra pars reticulata; SNc: substantia nigra pars compacta. Scale bar: 500 microns.
[0016] [Diagram 5] Summary of the results of the dopaminergic enhancer screening based on the analysis of virus-mediated detection of native SYFP2 signal in midbrain structures (SNc: substantia nigra pars compacta; VTA: ventral tegmental area) and immunohistochemical analysis by co-labeling with anti-TH and anti-GFP antibodies. Expression intensity is indicated as "strong" or "weak" based on the signal intensity of the reporter transgene. ND: not determined.
[0017] [Figure 6]The sequences supporting the present disclosure are shown below. These sequences include eHGT_606h (SEQ ID NO: 1), core4_eHGT_606h (SEQ ID NO: 2), 3xcore4_eHGT_606h (SEQ ID NO: 3), core5_eHGT_606h (SEQ ID NO: 4), 3xcore5_eHGT_606h (SEQ ID NO: 5), Core_hTH (SEQ ID NO: 6), 3xCore_hTH (SEQ ID NO: 7), Core_eHGT_1039h (SEQ ID NO: 8), 3xCore_eHGT_1039h (SEQ ID NO: 9), eHGT_888m (SEQ ID NO: 10), Core4_eHGT_888m (SEQ ID NO: 83), 3Xcore4_eHGT_888m (SEQ ID NO: 84), Core4_eHGT_888h (SEQ ID NO: 85), 3Xcore4_eHGT_888h (840bp) (SEQ ID NO: 86), eHGT_017m (SEQ ID NO: 87), core1_eHGT_017m (SEQ ID NO: 88), 3Xcore1_eHGT_017m (SEQ ID NO: 89), MGT_E51 (SEQ ID NO: 13), Core2-MGT_E51 (SEQ ID NO: 90), 3xCore2-MGT_E51 (SEQ ID NO: 91), eHGT_897m (SEQ ID NO: 11), MGT_E68 (SEQ ID NO: 12) , eHGT_889h (SEQ ID NO: 15), eHGT_1038m (SEQ ID NO: 16), eHGT_589m (SEQ ID NO: 92), eHGT_647m (SEQ ID NO: 93), eHGT_483m (SEQ ID NO: 94), β-globin minimal promoter (pBGmin / minBGlobin / minBGprom) (SEQ ID NO: 17), minCMV promoter (SEQ ID NO: 18), mutated minCMV promoter (removal of SacI restriction enzyme site) (SEQ ID NO: 19), minRho promoter (SEQ ID NO: 20), minRho* promoter (SEQ ID NO: 21), Hsp68 minimal promoter (proHsp68) (SEQ ID NO: 22), SYFP2 (SEQ ID NO: 23), EGFP (SEQ ID NO: 24), optimized Flp recombinase (FlpO) (SEQ ID NO: 25), improved Cre recombinase (iCre) (SEQ ID NO: 26), SP10 insulator (SP10ins) (SEQ ID NO: 27), 3xSP10ins (SEQ ID NO: 28), 3XFLAG (SEQ ID NO: 29), 10aa (SEQ ID NO: 30), H2B (SEQ ID NO: 31), H2B* (SEQ ID NO: 95), WPRE3 (SEQ ID NO: 32), WPRE (SEQ ID NO: 33),BGHpA (SEQ ID NO:34), hGHpA (SEQ ID NO:35), P2A (SEQ ID NO:36), T2A (SEQ ID NO:37), E2A (SEQ ID NO:38), F2A (SEQ ID NO:39), Exemplary Plasmid Backbone 1 - Left ITR (SEQ ID NO:40), Exemplary Plasmid Backbone 1 - Right ITR (SEQ ID NO:41), Exemplary Plasmid Backbone 2 - Left ITR (SEQ ID NO:42), Exemplary Plasmid Backbone 2 - Right ITR (SEQ ID NO:43), PHP.eB capsid (SEQ ID NO:44), AAV9 VP1 capsid protein (SEQ ID NO: 45), tet-transactivator version 2 (tTA2) (SEQ ID NO: 46), GTPase HRas [Homo sapiens] (SEQ ID NO: 47), substance P, which is the 58th to 68th residues of protachykinin-1 [Homo sapiens] (SEQ ID NO: 48), oxytocin, which is the 20th to 28th residues of oxytocin-neurophysin-1 [Homo sapiens] (SEQ ID NO: 49), GCaMP6m (SEQ ID NO: 50), GCaMP6s (SEQ ID NO: 51), GCaMP6f (SEQ ID NO: 52), CN3406 (SEQ ID NO: 53), CN3509 (SEQ ID NO: 54), No. 54), CN3058 (SEQ ID NO: 55), CN3059 (SEQ ID NO: 56), CN3829 (SEQ ID NO: 57), CN2436 (SEQ ID NO: 58), CN3283 (SEQ ID NO: 59), AiP1240 (SEQ ID NO: 60), CN3739 (SEQ ID NO: 63), CN3889 (SEQ ID NO: 64), CN2839 (SEQ ID NO: 96), CN2847 (SEQ ID NO: 97), CN2431 (SEQ ID NO: 98), CN3863 (SEQ ID NO: 99), CN4365 (SEQ ID NO: 100), CN4830 (SEQ ID NO: 101), CN3551 (SEQ ID NO: 102), and AiP1417 (SEQ ID NO: 103). DETAILED DESCRIPTION OF THE PREFERRED EMBODIMENTS
[0018] To fully understand the biology of the brain, we need to distinguish, define, and dissect different cell types, as well as identify artificial expression constructs that can label and perturb them (Tasic, Curr. Opin. Neurobiol. 50, 242-249 (2018); Zeng & Sanes, Nat. Rev. Neurosci. 18, 530-546 (2017)). In mice, driver lines expressing recombinases have been used successfully to label cell populations that share marker gene expression (Daigle et al., Cell 174, 465-480.e22 (2018); Taniguchi, et al., Neuron 71, 995-1013 (2011); Gong et al., J. Neurosci. 27, 9817-9823 (2007)). However, the creation, maintenance and use of such lines capable of labeling specific cell types with high specificity is costly and often requires crossing of transgenic animals between three species, resulting in low frequency of obtaining the desired experimental animals. Moreover, these tools cannot be applied to humans because of the need for germline transgenic animals.
[0019] The present disclosure provides artificial expression constructs that direct gene expression in dopaminergic neurons of the central nervous system.
[0020] In certain embodiments of the artificial expression constructs of the present disclosure, an enhancer selected from eHGT_888m, eHGT_897m, eHGT_606h, MGT_E68, eHGT_889h, eHGT_1038m, eHGT_589m, eHGT_647m and eHGT_483m is utilized to induce gene expression in dopaminergic neurons.
[0021] In certain embodiments, these artificial enhancer elements are concatemerized enhancer cores.Examples of such enhancer cores include MGT_E51, eHGT_606h, eHGT_1039h, hTH, eHGT_888m, eHGT_888h and / or eHGT_017m cores, or concatemerized these cores.By using such artificial enhancer elements, it is possible to rapidly express transgenes and obtain high expression, compared to using the full length of original (natural) enhancer alone.
[0022] In certain embodiments, the enhancer cores include sequences shown in SEQ ID NO:2, SEQ ID NO:4, SEQ ID NO:6, SEQ ID NO:8, SEQ ID NO:83, SEQ ID NO:85, SEQ ID NO:90, or SEQ ID NO:88. In certain embodiments, these enhancer cores are concatemerized and have 2, 3, 4, 5, 6, 7, 8, 9, or 10 copies of the core sequence. In certain embodiments, concatemers containing 3 copies of the selected enhancer core include sequences shown in any of SEQ ID NO:3, SEQ ID NO:5, SEQ ID NO:7, SEQ ID NO:9, SEQ ID NO:84, SEQ ID NO:86, SEQ ID NO:91, and SEQ ID NO:89.
[0023] Certain embodiments of the artificial expression constructs of the present disclosure utilize 3xCore-MGT_E51, 3xcore4_eHGT_606h, 3xcore5_eHGT_606h, 3xcore_eHGT_1039h, 3xCore_hTH, core4_eHGT_888m, 3Xcore4_eHGT_888m, 3Xcore4_eHGT_888h and / or 3xcore1_eHGT_017m to selectively induce gene expression in dopaminergic neurons. Typically, the cores are derived from enhancers, although this property is not required for all embodiments.
[0024] In certain embodiments, artificial expression constructs are provided that include features of vectors described herein, such as vectors selected from CN3406, CN3509, AiP1417, CN2436, AiP1240, CN3739, CN3889, CN2839, CN2847, CN2431, CN3058, CN3059, CN3829, CN3283, CN3863, CN4365, CN4830, and CN3551.
[0025] Various aspects of the disclosure are described in more detail below with further options. Various aspects of the disclosure are described under the following headings: (i) artificial expression constructs and vectors for targeted expression of genes in targeted cell types; (ii) compositions for administration; (iii) cell lines containing artificial expression constructs; (iv) transgenic animals; (v) methods of use; (vi) kits and commercial packages; (vii) exemplary embodiments; and (viii) conclusion. These headings are provided for organizational purposes only and are not intended to limit the scope or interpretation of the disclosure.
[0026] (i) Artificial expression constructs and vectors for targeted expression of genes in target cell types. The artificial expression constructs disclosed herein comprise (i) an enhancer sequence that induces targeted expression of a coding sequence in a targeted type of central nervous system cell, (ii) the coding sequence to be expressed, and (iii) a minimal promoter. The artificial expression constructs of the invention may further comprise other regulatory elements as needed or beneficial.
[0027] In certain embodiments, an "enhancer" or "enhancer element" is a cis-acting sequence that increases the amount of transcription associated with a promoter, can function in either the forward or reverse orientation relative to the promoter and the coding sequence to be transcribed, and can be located upstream or downstream relative to the promoter or the coding sequence to be transcribed. There are various methods or techniques known in the art for measuring the function of enhancer element sequences. Specific examples of enhancer sequences used in the artificial expression constructs disclosed herein include eHGT_888m, eHGT_897m, eHGT_606h, MGT_E68, eHGT_889h, eHGT_1038m, eHGT_589m, eHGT_647m and eHGT_483m; and concatemerized enhancer cores such as 3xCore2-MGT_E51, 3xcore4_eHGT_606h, 3xcore5_eHGT_606h, 3xcore_eHGT_1039h, 3xCore_hTH, core4_eHGT_888m, 3Xcore4_eHGT_888m, 3Xcore4_eHGT_888h, 3xcore1_eHGT_017m.
[0028] In certain embodiments, the enhancer used in the target type of central nervous system cells is the enhancer that is only utilized in the target type of central nervous system cells, or the enhancer that is mainly utilized in the target type of central nervous system cells.The enhancer used in the target type of central nervous system cells is the enhancer that enhances the expression of genes in the target type of central nervous system.In certain embodiments, the enhancer used in the target type of central nervous system cells enhances the expression of genes in the target type of central nervous system, but does not substantially induce the expression of genes in other non-target cells, and therefore is also a selective target central nervous system enhancer that has cell type-specific transcription activity.
[0029] When a heterologous coding sequence operably linked to an enhancer disclosed herein is expressed in a cell type that is targeted, the administered heterologous coding sequence is expressed in the intended cell type.
[0030] When a heterologous coding sequence is preferentially expressed in selected cells, the administered heterologous coding sequence is expressed in the intended cell type, but is not substantially expressed in other cell types. This is described in more detail below. In certain embodiments, not substantially expressed in other cell types means less than 50% expression in the reference cell compared to the target cell type; less than 40% expression in the reference cell compared to the target cell type; less than 30% expression in the reference cell compared to the target cell type; less than 20% expression in the reference cell compared to the target cell type; or less than 10% expression in the reference cell compared to the target cell type. In certain embodiments, "reference cell" refers to a non-target cell. The non-target cell may be in the same anatomical structure as the target cell and / or may project to a common anatomical region. In certain embodiments, the reference cell is in an anatomical structure adjacent to the anatomical structure containing the target cell type. In certain embodiments, the reference cell is a non-target cell that has a different gene expression profile than the target cell.
[0031] In certain embodiments, the transcription product of the coding sequence may be expressed at a low level in unselected cell types, for example, at less than 1% or 1%, 2%, 3%, 5%, 10%, 15% or 20% of the transcription product expression in selected cells. In certain embodiments, the target type of central nervous system cell is the only type of cell that can express the correct combination of various transcription factors that can bind to the enhancer disclosed herein and induce gene expression. Thus, in certain embodiments, expression occurs only in the target type of cell.
[0032] In certain embodiments, target cell types (e.g., neuronal and / or non-neuronal cells) can be identified based on transcriptional profiles, such as those described in Tasic et al., Nature 563, 72-78 (2018) and Hodge et al., Nature 573, 61-68 (2019).
[0033] Dopaminergic neurons secrete dopamine. Midbrain dopaminergic neurons are the main source of dopamine in the mammalian central nervous system. Loss of dopaminergic neurons is associated with Parkinson's disease, one of the most well-known human neurological disorders. Studies on the developmental pathways involved in the generation of dopaminergic neurons in the brain have identified several specific transcription factors, such as Nurr1, Lmx1b, and Pitx3, all of which have been shown to be important in the development of the midbrain dopaminergic system.
[0034] In certain embodiments, the coding sequence is a heterologous coding sequence that codes for an effector element.Effector element is the sequence that is expressed to obtain an intended effect, and the intended effect is actually achieved by this effector element.Examples of effector elements include reporter genes / proteins and functional genes / proteins.
[0035] Exemplary reporter genes / proteins include those expressed by Addgene ID No. 83894 (pAAV-hDlx-Flex-dTomato-Fishell_7), ID No. 83895 (pAAV-hDlx-Flex-GFP-Fishell_6), ID No. 83896 (pAAV-hDlx-GiDREADD-dTomato-Fishell-5), ID No. 83898 (pAAV-mDlx-ChR2-mCherry-Fishell-3), ID No. 83899 (pAAV-mDlx-GCaMP6f-Fishell-2), ID No. 83900 (pAAV-mDlx-GFP-Fishell-1), and ID No. 89897 (pcDNA3-FLAG-mTET2(N500)). Exemplary reporter genes include, inter alia, expressible fluorescent proteins or expressible biotin; blue fluorescent proteins (e.g., eBFP, eBFP2, Azurite, mKalama1, GFPuv, Sapphire, T-sapphire); cyan fluorescent proteins (e.g., eCFP, Cerulean, CyPet, AmCyanl, Midoriishi-Cyan, mTurquoise); green fluorescent proteins (e.g., GFP, GFP-2, tagGFP, turboGFP, EGFP, Emerald, Azami Green, Monomeric Azami Green (mAzamigreen), CopGFP, AceGFP, avGFP, ZsGreen1, Oregon Green, TM (Thermo Fisher Scientific); luciferase; orange fluorescent proteins (mOrange, mKO, Kusabira-Orange, Monomeric Kusabira-Orange, mTangerine, tdTomato, dTomato); red fluorescent proteins (mKate, mKate2, mPlum, DsRed monomer, mCherry, mRuby, mRFP1, DsRed-Express, DsRed2, DsRed-Monomer, HcRed-Tandem, HcRedl, AsRed2, eqFP611, mRaspberry, mStrawberry, Jred, Texas Red) TM(Thermo Fisher Scientific)); far-red fluorescent proteins (e.g., mPlum and mNeptune); yellow fluorescent proteins (e.g., YFP, eYFP, Citrine, SYFP2, Venus, YPet, PhiYFP, ZsYellow1); or reporter genes encoding tandemly linked complexes.
[0036] GFP is composed of 238 amino acids (26.9 kDa) and was first isolated from the jellyfish Aequorea victoria / Aequorea aequorea / Aequorea forskalea, which fluoresces green when exposed to blue light. GFP isolated from A. victoria has a major excitation peak at a wavelength of 395 nm and a minor excitation peak at 475 nm. Its emission peak is at 509 nm, which is in the low wavelength range of green light in the visible spectrum. GFP from the sea pansy (Renilla reniformis) has one major excitation peak at 498 nm. Due to its wide range of applications and the demand from researchers for further improvements, various GFP variants have been created. The first major improvement was a single point mutation (S65T) reported in Nature by Roger Tsien in 1995. This mutation dramatically improved the spectral properties of GFP, increasing its fluorescence and photostability, and shifting the major excitation peak to 488 nm while maintaining the emission peak at 509 nm. Enhanced GFP (EGFP) was obtained by adding a point mutation (F64L) to GFP that improves folding efficiency at 37 °C. EGFP has an extinction coefficient (denoted ε) of 55,000 L / mol cm, which is 9.13 × 10 per molecule. -21 m 2 Also known as the optical cross section of GFP, Superfolder GFP was reported in 2006 as a series of GFP mutants that can rapidly fold and mature even when fused to poorly folded peptides.
[0037] "Yellow fluorescent protein" (YFP) is a genetic variant of the green fluorescent protein derived from Aequorea victoria. Its excitation peak is at 514 nm and its emission peak is at 527 nm.
[0038] Exemplary functional molecules include ion transporters, cellular trafficking proteins, enzymes, transcription factors, neurotransmitters, calcium reporters, channelrhodopsins, guide RNAs, nucleases, microRNAs, or designer receptors activated only by designer drugs (DREADDs), each of which has a function.
[0039] Ion transporters are transmembrane proteins responsible for the transport of ions across cell membranes. Ion transporters are found in most cell types and are important in regulating cellular excitability and homeostasis. Ion transporters are involved in numerous cellular processes, including action potentials, synaptic transmission, hormone secretion, and muscle contraction. Many biological processes important to living cells involve the transport of calcium ions (Ca) through ion channels. 2+ ), potassium ion (K + ), sodium ion (Na + In certain embodiments, ion transporters include voltage-gated sodium channels (e.g., SCN1A), potassium channels (e.g., KCNQ2), and calcium channels (e.g., CACNA1C).
[0040] Exemplary enzymes, transcription factors, receptors, membrane proteins, cellular transport proteins, signaling molecules and neurotransmitters include enzymes such as lactase, lipase, helicase, α-glucosidase, amylase, aromatic L-amino acid decarboxylase (AADC); transcription factors such as SP1, AP-1, heat shock factor protein 1, C / EBP (CCAAT / enhancer binding protein), Oct-1; transforming growth factor receptor β1, platelet-derived growth factor receptor, epidermal growth factor receptor, vascular endothelial growth factor receptor, interleukin-1 (IL-1), and IL-2; These include receptors such as leukin-8 receptor α; membrane proteins and cell trafficking proteins such as clathrin, dynamin, caveolin, Rab4A, and Rab-11A; signaling molecules such as nerve growth factor (NGF), glial cell line-derived neurotrophic factor (GDNF), platelet-derived growth factor (PDGF), transforming growth factor beta (TGFβ), epidermal growth factor (EGF), GTPases, and HRas; and neurotransmitters such as cocaine- and amphetamine-regulated transcript, substance P, oxytocin, and somatostatin.
[0041] In certain embodiments, the functional molecules include reporters that indicate cell function and state, such as calcium reporters. Intracellular calcium concentration is an important predictor of many cellular activities, such as neuronal activation, muscle cell contraction, and second messenger signaling. A highly sensitive and simple technique for monitoring intracellular calcium concentration is the use of genetically encoded calcium indicators (GECIs). Among GECIs, a green fluorescent protein (GFP)-based calcium sensor, named GCaMP, is highly efficient and widely used. GCaMP is formed by fusing M13 and calmodulin proteins to the N- and C-termini of circularly permuted GFP. Some types of GCaMP exhibit characteristic fluorescence emission spectra (Zhao et al., Science, 2011, 333(6051): 1888-1891). Exemplary GECIs that exhibit green fluorescence include GCaMP3, GCaMP5G, GCaMP6s, GCaMP6m, GCaMP6f, jGCaMP7s, jGCaMP7c, jGCaMP7b, jGCaMP7f, jGCaMP8s, jGCaMP8m and jGCaMP8f.Furthermore, GECIs that exhibit red fluorescence include jRGECO1a and jRGECO1b.AAV products that contain GECIs are commercially available.For example, AAV8-CAG-GCaMP3 (catalog no. BS4-CX3AAV8), AAV8-Syn-FLEX-GCaMP6s-WPRE (catalog no. BS1-NXSAAV8), AAV8-Syn-FLEX-GCaMP6s-WPRE (catalog no. BS1-NXSAAV8), AAV9-CAG-FLEX-GCaMP6m-WPRE (catalog no. BS2-CXMAAV9), AAV9-Syn-FLEX-jGCaMP7s-WPRE ( AAV products available include AAV9-CAG-FLEX-jGCaMP7f-WPRE (Catalog No. BS12-CXFAAV9), AAV9-Syn-FLEX-jGCaMP7b-WPRE (Catalog No. BS12-NXBAAV9), AAV9-Syn-FLEX-jGCaMP7c-WPRE (Catalog No. BS12-NXCAAV9), AAV9-Syn-FLEX-NES-jRGECO1a-WPRE (Catalog No. BS8-NXAAAV9), and AAV8-Syn-FLEX-NES-jRCaMP1b-WPRE (Catalog No. BS7-NXBAAV8).
[0042] In certain embodiments, the calcium reporter includes the genetically encoded calcium indicator (GECI) NTnC; a myosin light chain kinase-GFP-calmodulin chimera; the calcium indicator TN-XXL; a BRET-based auto-luminescent calcium indicator; and / or the calcium indicator protein OeNL(Ca2+)-18μ.
[0043] In certain embodiments, the functional molecules include modulators of neuron-active channelrhodopsins (e.g., channelrhodopsin 1, channelrhodopsin 2, and variants thereof). Channelrhodopsins are a subfamily of retinylidene proteins (rhodopsins) that function as light-gated ion channels. In addition to channelrhodopsin 1 (ChR1) and channelrhodopsin 2 (ChR2), several channelrhodopsin variants have been developed. For example, Lin et al. (Biophys J, 2009, 96(5): 1803-14) describe the creation of transmembrane domain chimeras of ChR1 and ChR2 using site-directed mutagenesis. Zhang et al. (Nat Neurosci, 2008, 11(6): 631-3) describe a red-light shifted channelrhodopsin variant, VChR1. VChR1 has reduced photosensitivity and reduced membrane trafficking and expression. Other known channelrhodopsin variants include the ChR2 variant described in Nagel, et al., Proc Natl Acad Sci USA, 2003, 100(24): 13940-5, ChR2 / H134R (Nagel, G., et al., Curr Biol, 2005, 15(24): 2279-84) and ChD / ChEF / ChIEF (Lin, JY, et al., Biophys J, 2009, 96(5): 1803-14), all of which are activated by blue light (470 nm) but are insensitive to orange / red light. Other variants are described in Lin, Experimental Physiology, 2010, 96.1: 19-25; Knopfel et al., The Journal of Neuroscience, 2010, 30(45): 14998-15004; and Mardinly et al., Nat Neurosci. 2018, 21(6):881-893.
[0044] In certain embodiments, the functional molecule includes DNA and RNA editing tools, such as CRISPR / Cas (e.g., guide RNA and nuclease such as Cas, Cas9, cpf1).Furthermore, the functional molecule includes recombinant Cpf1 as described in US Patent Publication No. 2018 / 0030425, US Patent Publication No. 2016 / 0208243, WO / 2017 / 184768 and Zetsche et al. (2015) Cell 163: 759-771; single-stranded gRNA (see, for example, Jinek et al. (2012) Science 337:816-821; Jinek et al. (2013) eLife 2:e00471; Segal (2013) eLife 2:e00563), editase, guide RNA molecule, microRNA, or homologous recombination donor cassette.
[0045] In certain embodiments, the functional molecule includes a localization cassette. In certain embodiments, the localization cassette is used to localize a molecule (e.g., a vector, a protein, a sensor) to a specific subcellular compartment, such as the cell body, axon, or dendrite of a neuron. In certain embodiments, the localization cassette includes a somatic tag (e.g., soma (EE-RR)) for localization to the soma; an axon tag (e.g., derived from GAP43) or synaptophysin (sy) for localization to the axon; a hydrophobic tail for localization to the cell membrane; and a hydrophobic or alkyl chain for localization to the endoplasmic reticulum. In certain embodiments, the localization cassette is fused to a sensor molecule, such as a GECI. In certain embodiments, the fusion protein of the localization cassette and the GECI includes soma-jGCaMP8s, axon-jRGECO1a, syGCaMP5G, and soma-jGCaMP7s.
[0046] In a particular embodiment, the functional molecule includes a tag cassette. Examples of the tag cassette include His tag (HHHHHH; SEQ ID NO: 65), Flag tag (DYKDDDDK; SEQ ID NO: 66), Xpress tag (DLYDDDDK; SEQ ID NO: 67), Avi tag (GLNDIFEAQKIEWHE; SEQ ID NO: 68), calmodulin tag (KRRWKKNFIAVSAANRFKKISSSGAL; SEQ ID NO: 69), polyglutamic acid tag, HA tag (YPYDVPDYA; SEQ ID NO: 70), Myc tag (EQKLISEEDL; SEQ ID NO: 71), Strep tag (meaning the original STREP tag) (WRHPQFGG; SEQ ID NO: 72), STREP tag II (WSHPQFEK; SEQ ID NO: 73 (Institut fur Bioanalytik (IBA) GmbH, Germany; see, for example, U.S. Patent Publication No. 7,981,632), Softag 1 (SLAELLNAGLGGS; SEQ ID NO: 74), Softag 3 (TQDPSRVG; SEQ ID NO: 75), and V5 tag (GKPIPNPLLGLDST; SEQ ID NO: 76). In certain embodiments, the tag cassette includes a fusion tag cassette such as 3XFLAG. In certain embodiments, the 3XFLAG includes the sequence shown in SEQ ID NO: 29.
[0047] The sequences of the aforementioned functional molecules have been published, for example, lactase (e.g., GenBank: EAX11622.1), lipase (e.g., GenBank: AAA60129.1), helicase (e.g., GenBank: AMD82207.1), amylase (e.g., GenBank: AAA51724.1), α-glucosidase (e.g., GenBank: ABI53718.1), transcription factor SP1 (e.g., UniProtKB / Swiss-Prot: P08047.3), transcription factor AP-1 (e.g., NP_002219.1), heat shock factor protein 1 (e.g., UniProtK B / Swiss-Prot:Q00613.1), CCAAT / enhancer-binding protein (C / EBP) beta isoform a (e.g. NP_005185.2), Oct-1 (e.g. UniProtKB / Swiss-Prot:P14859.2), TGF-β (e.g. GenBank:CAF02096.2), glial cell line-derived neurotrophic factor (GDNF) (e.g. NP_001177397.1), platelet-derived growth factor receptor (e.g. GenBank:AAA60049.1), epidermal growth factor receptor (e.g. GenBank:CAA25 240.1), vascular endothelial growth factor receptor (e.g., GenBank: AAC16449.2), interleukin-8 receptor α (e.g., GenBank: AAB59436.1), caveolin (e.g., GenBank: CAA79476.1), dynamin (e.g., GenBank: AAA88025.1), clathrin heavy chain 1 isoform 1 (e.g., NP_004850.1), clathrin heavy chain 2 isoform 1 (e.g., NP_009029.3), clathrin light chain A isoform a (e.g., NP_001824.1), clathrin light chain B isoform a (e.g. NP_001825.1), ras-related protein Rab-4A isoform 1 (e.g. NP_004569.2), ras-related protein Rab-11A (e.g. UniProtKB / Swiss-Prot:P62491.3), platelet-derived growth factor (e.g. GenBank:AAA60552.1), transforming growth factor beta 3 (e.g. GenBank:AAA61161.1), nerve growth factor (e.g. GenBank:CAA37703.1), EGF (e.g. GenBank:CAA34902.2), cocaine-amphetamine regulated transcript (A chain) (e.g. PDB:1HY9_A), protachykinin-1 (e.g. UniProtKB-P20366), oxytocin neurophysin 1 (e.g. UniProtKB-P01178), somatostatin (e.g. GenBank:AAH32625.1), genetically encoded green calcium indicator NTnC (A chain) [synthetic construct] (e.g. PDB:5MWC_A), calcium indicator TN-XXL [synthetic construct] (e.g. GenBank:ACF93133.1), BRET-based auto-luminescent calcium indicator [synthetic construct] (e.g., GenBank: ADF42668.1), calcium indicator protein OeNL(Ca2+)-18μ [synthetic construct] (e.g., GenBank: BBB18812.1), myosin light chain kinase, green fluorescent protein, calmodulin chimera (A chain) [synthetic construct] (e.g., PDB: 3EKJ_A), channelopsin 1 (e.g., UniProtKB-F8UVI5), channelopsin 1 (e.g., GenBank: AER58217.1), channelrhodopsin 2 (e.g., UniProtKB-B4Y105), channelrhodopsin 2 [synthetic construct] (e.g., GenBank: ABO64386.1), CRISPR-associated proteins (Cas) (e.g., GenBank: AKG27598.1), Cas9 [synthetic construct] (e.g., GenBank: AST09977.1), CRISPR-associated endonuclease Cpf1 (e.g., UniProtKB / Swiss-Prot: U2UMQ6.1), ribonuclease Examples of such proteins include ribonuclease 4 or ribonuclease L (e.g. UniProtKB / Swiss-Prot:Q05823.2), deoxyribonuclease IIβ (e.g. GenBank:AAF76893.1), sodium channel protein type 1 subunit α (e.g. UniProtKB-P35498), member 2 of the voltage-gated potassium channel subfamily KQT (e.g. UniProtKB-O43526) and voltage-gated L-type calcium channel subunit α-1C (e.g. UniProtKB-Q13936).
[0048] Further effector elements include Cre, iCre, dgCre, FlpO and tTA2. iCre refers to codon-improved Cre. dgCre is a GFP / Cre recombinase fusion gene enhanced by the N-terminal fusion of the first 159 amino acids of the dihydrofolate reductase gene (DHFR or folA) of the E. coli K12 chromosome, which has a G67S mutation and a destabilization domain mutation R12Y / Y100I upon recombination. FlpO is a codon-optimized form of FLPe, which significantly improves protein expression and FRT recombination efficiency in mouse cells. The FLP / FRT system is widely used for gene expression, as is the Cre / LoxP system (the generation of conditional knockout mice using the FLP / FRT system is also widely practiced). tTA2 refers to tetracycline transactivator.
[0049] Exemplary expressible elements include expression products that do not include effector elements, such as non-functional or defective proteins. In certain embodiments, such expressible elements can be used to perform methods for testing the effect of their corresponding functional molecules. In certain embodiments, the expressible elements are non-functional or defective due to recombinant mutations that disable their function. In these aspects, the non-expressible elements are as similar in structure as possible to their corresponding functional molecules.
[0050] Exemplary self-cleaving peptides include 2A peptides, which allow two proteins to be produced from one mRNA. 2A sequences are short sequences (e.g., 20 amino acids long) and are often used in size-restricted constructs. Specific examples include P2A, T2A, E2A, and F2A. In certain embodiments, the artificial expression construct comprises an internal ribosome entry site (IRES) sequence. The IRES can initiate ribosome translation from a second internal site on the mRNA molecule, allowing two proteins to be produced from one mRNA.
[0051] The artificial expression construct may encode nuclear transport proteins such as histone H1, histone H2A, histone H2B, histone H3, histone H4, and the histone-like proteins HPhA and H2B*.
[0052] Coding sequences encoding the molecules (e.g., RNA and proteins) described herein can be obtained from publicly available databases and publications. The coding sequences may further contain various sequence polymorphisms, mutations and / or sequence variants, and such changes do not affect the function of the encoded molecule. "Encode" refers to the property of a nucleic acid sequence, such as a vector, plasmid, gene, cDNA, mRNA, etc., to function as a template for the synthesis of other molecules, such as proteins.
[0053] The term "gene" may include not only coding sequences, but also regulatory regions such as promoters, enhancers, insulators and / or post-transcriptional regulatory elements (e.g., termination regions). In addition, the term may include any introns and other DNA sequences spliced from the mRNA transcript, as well as variants resulting from alternative splice sites. These sequences may further include sequences or degenerate codons of a reference sequence that may be introduced to confer codon preference in a particular type of organism or cell.
[0054] The promoter may be a general promoter, a tissue-specific promoter, a cell-specific promoter, and / or a cytoplasm-specific promoter. The promoter may be a strong promoter, a weak promoter, a constitutive expression promoter, and / or an inducible promoter. An inducible promoter induces expression in response to a specific condition, signal, or cellular event. For example, the promoter may be an inducible promoter that requires a specific ligand, small molecule, transcription factor, or hormone protein to induce transcription from the promoter. Specific examples of promoters include minBglobin (also called minBGprom), CMV, minCMV, minCMV* (minCMV* is minCMV with the SacI restriction site removed), minRho, minRho* (minRho* is minRho with the SacI restriction site removed), SV40 immediate early promoter, Hsp68 minimal promoter (proHSP68), and Rous sarcoma virus (RSV) long terminal repeat (LTR) promoter. A minimal promoter does not have the activity of inducing gene expression by itself, but when linked to an enhancer element nearby, it is activated and can induce gene expression.
[0055] In certain embodiments, the expression construct is provided in a vector. A "vector" refers to a nucleic acid molecule capable of transferring or transporting another nucleic acid molecule, such as an expression construct. The transferred nucleic acid is usually linked to, e.g., inserted into, the nucleic acid molecule of the vector. The vector may contain a sequence that induces autonomous replication of the cell or may contain a sequence that allows integration into the DNA of the host cell. Useful vectors include, for example, plasmids (e.g., DNA and RNA plasmids), transposons, cosmids, bacterial artificial chromosomes, and viral vectors.
[0056] The term "viral vector" is used broadly to refer to a nucleic acid molecule that contains components derived from a virus that facilitate the transfer and expression of non-natural nucleic acid molecules in cells. An "adeno-associated viral vector" refers to a viral vector or plasmid that contains structural and functional genetic elements or portions thereof that are primarily derived from AAV. A "retroviral vector" refers to a viral vector or plasmid that contains structural and functional genetic elements or portions thereof that are primarily derived from a retrovirus. A "lentiviral vector" refers to a viral vector or plasmid that contains structural and functional genetic elements or portions thereof that are primarily derived from a lentivirus or the like. A "hybrid vector" refers to a vector that contains structural and / or functional genetic elements that are primarily derived from more than one virus.
[0057] "Adenoviral vector" refers to a construct that contains sufficient adenoviral sequences (a) to facilitate packaging of an artificial expression construct, and (b) to express a coding sequence cloned in the sense or antisense orientation. Recombinant adenoviral vectors include genetically engineered forms of adenovirus. The genetic makeup of adenovirus is a 36 kb linear double-stranded DNA virus, allowing replacement of large portions of adenoviral DNA with up to 7 kb of foreign sequence. Adenoviral DNA can replicate episomally without causing genotoxicity, and thus, unlike retroviruses, is not integrated into chromosomes upon infection of a host cell with adenovirus. Furthermore, adenoviruses are structurally stable, and no genome rearrangements have been detected after extensive amplification.
[0058] Adenoviruses are particularly suitable for use as gene transfer vectors due to their moderate genome size, ease of manipulation, high titer, wide target cell range and high infectivity. Both ends of the adenovirus genome contain inverted repeats (ITRs) of 100-200 base pairs in length, which are cis elements required for viral DNA replication and packaging. The early (E) and late (L) regions of the adenovirus genome contain various transcription units that are divided by the initiation of viral DNA replication. The E1 region (E1A and E1B) encodes proteins responsible for regulating the transcription of the adenoviral genome and several cellular genes. Expression of the E2 region (E2A and E2B) results in the synthesis of proteins for viral DNA replication. These proteins are involved in DNA replication, expression of late genes and shut-off of host cell protein biosynthesis. Late gene products, including the majority of adenovirus capsid proteins, are expressed only after significant processing of a single primary transcript driven by the major late promoter (MLP). The MLP is particularly efficient at the late stages of infection, when all mRNAs driven by this promoter have a tripartite 5'-leader (TPL) sequence and are preferentially selected by the mRNA for translation.
[0059] Other than the requirement that the adenoviral vector be replication-deficient or at least conditionally-deficient, the characteristics of the adenoviral vector are not believed to be critical to the successful implementation of certain embodiments disclosed herein. The adenovirus may be of any of the 42 known serotypes or subgenuses A-F. In certain embodiments, adenovirus of serotype 5 of the C subgenus is preferred as the starting material for obtaining a conditionally replication-deficient adenoviral vector for use in certain embodiments, since type 5 adenovirus is a human adenovirus with a large amount of known biochemical and genetic information and has been used historically in the majority of constructions using adenoviruses as vectors.
[0060] As described herein, the vector is generally replication-defective and lacks the E1 region of adenovirus. Therefore, it is most convenient to introduce the polynucleotide encoding the gene of interest into the position where the coding sequence of the E1 region has been deleted. However, the insertion position of the construct within the adenovirus sequence is not critical. The polynucleotide encoding the gene of interest may be inserted into the deleted E3 region of the E3 replacement vector, or into the E4 region, and the defect in the E4 region is complemented by a helper cell line or a helper virus.
[0061] Adeno-associated virus (AAV) is a parvovirus found as a contaminant of adenovirus stocks. AAV is a ubiquitous virus (85% of the US population has anti-AAV antibodies) and does not cause disease. Furthermore, AAV is classified as a dependovirus because its replication depends on the presence of a helper virus (e.g., adenovirus). Various serotypes have been isolated, of which AAV-2 is the most extensively characterized. AAV has a single-stranded linear DNA that is packaged with the capsid proteins VP1, VP2, and VP3 to form icosahedral virions with a diameter of 20-24 nm.
[0062] The length of AAV DNA is 4.7 kilobases. AAV DNA contains two open reading frames, flanked by two ITRs. There are two main genes in the AAV genome: rep and cap. The rep gene codes for the proteins responsible for AAV viral replication, and the cap gene codes for the capsid proteins VP1-3. Each ITR forms a T-shaped hairpin structure. These terminal repeats are the only cis components of AAV required for chromosomal integration. Thus, AAV can be used as a vector in which all viral coding sequences can be removed and replaced with gene cassettes for delivery. Three AAV viral promoters have been identified and named p5, p19, and p40, respectively, based on their map location. Transcription from p5 and p19 results in the production of the rep protein, and transcription from p40 produces the capsid protein.
[0063] AAV is outstanding for use in the present disclosure because it has a good safety profile and can be expressed in target cell populations by modifying capsid and genome.scAAV refers to self-complementary AAV.pAAV refers to plasmid adeno-associated virus.rAAV refers to recombinant adeno-associated virus.
[0064] Other viral vectors may also be used, such as those derived from viruses such as vaccinia virus, poliovirus, herpes virus, etc. These vectors offer beneficial characteristics for a variety of mammalian cells.
[0065] Retroviruses are commonly used tools for gene delivery. A retrovirus is an RNA virus whose genomic RNA is reverse transcribed to produce a double-stranded linear DNA copy, which is then covalently integrated into the host genome. Once integrated into the host genome, the retrovirus is called a provirus. The provirus functions as a template for RNA polymerase II to induce the expression of RNA molecules that code for structural proteins and enzymes required for the production of new viral particles.
[0066] Examples of retroviruses suitable for use in certain embodiments include Moloney murine leukemia virus (M-MuLV), Moloney murine sarcoma virus (MoMSV), Harvey murine sarcoma virus (HaMuSV), mouse mammary tumor virus (MuMTV), gibbon ape leukemia virus (GaLV), feline leukemia virus (FLV), spumavirus, Friend murine leukemia virus, murine stem cell virus (MSCV), Rous sarcoma virus (RSV), and lentiviruses.
[0067] "Lentivirus" refers to the complex retrovirus group (or complex retrovirus genus). Examples of lentiviruses include HIV (human immunodeficiency virus; including HIV type 1 and HIV type 2); Visna-Maedi virus (VMV); Caprine arthritis-encephalomyelitis virus (CAEV); Equine infectious anemia virus (EIAV); Feline immunodeficiency virus (FIV); Bovine immunodeficiency virus (BIV); and Simian immunodeficiency virus (SIV). In certain embodiments, a vector backbone based on HIV (i.e., HIV cis-acting sequence elements) can be used.
[0068] In some types of vectors, safety can be improved by replacing the U3 region of the 5'LTR, which induces the transcription of the viral genome in the production of viral particles, with a heterologous promoter. Examples of heterologous promoters that can be used for this purpose include, for example, Simian Virus 40 (SV40) (e.g., early or late) promoter, Cytomegalovirus (CMV) (e.g., immediate early) promoter, Moloney Murine Leukemia Virus (MoMLV) promoter, Rous Sarcoma Virus (RSV) promoter, and Herpes Simplex Virus (HSV) (thymidine kinase) promoter. Conventional promoters can induce high levels of transcription independent of Tat. Replacement of the U3 region with a heterologous promoter removes the complete U3 sequence from the virus production system, reducing the possibility of recombination that produces a replicable virus. In certain embodiments, a heterologous promoter has the additional advantage of being able to control the way the viral genome is transcribed. For example, the heterologous promoter can be an inducible promoter, such that the entire viral genome or a portion thereof is transcribed only when an inducer is present. Inducers include one or more compounds or physiological conditions, such as the culture temperature or pH of the host cell.
[0069] In certain embodiments, the viral vector comprises a TAR element. "TAR" refers to the "transactivation response" gene element present in the R region of the LTR of lentivirus. This element interacts with the transactivator (tat) gene element of lentivirus to enhance viral replication. However, this element is not required in the embodiment that replaces the U3 region of 5'LTR with a heterologous promoter.
[0070] The "R region" refers to the region in the retroviral LTR from the start of the cap site (i.e., the transcription start site) to just before the start of the poly(A) tail. The R region is also defined as the region between the U3 and U5 regions. The R region plays a role in moving nascent DNA from one end of the genome to the other during reverse transcription.
[0071] In certain embodiments, the expression of heterologous sequences in viral vectors can be increased by incorporating post-transcriptional regulatory elements and efficient polyadenylation sites into the viral vector, and a transcription termination signal may also be incorporated into the viral vector. Various post-transcriptional regulatory elements can increase the expression of heterologous nucleic acids. Examples of post-transcriptional regulatory elements include the Woodchuck Hepatitis Virus post-transcriptional regulatory element (WPRE; Zufferey et al., 1999, J. Virol., 73:2886); the Hepatitis B virus post-transcriptional regulatory element (HPRE) (Smith et al., Nucleic Acids Res. 26(21):4818-4827, 1998); and other post-transcriptional regulatory elements (Liu et al., 1995, Genes Dev., 9:1766). In certain embodiments, the vector includes a post-transcriptional regulatory element such as a WPRE or HPRE. In certain embodiments, the vector lacks or does not include a post-transcriptional regulatory element such as a WPRE or HPRE.
[0072] Expression of heterologous genes can be increased by elements capable of inducing efficient transcription termination and polyadenylation of heterologous nucleic acid transcripts. Transcription termination signals are usually found downstream of polyadenylation signals. In certain embodiments, vectors contain a polyadenylation signal at the 3' end of a polynucleotide encoding an expressed molecule (e.g., a protein). "Poly(A) site" or "poly(A) sequence" refers to a DNA sequence that induces both transcription termination and polyadenylation of a nascent RNA transcript transcribed by RNA polymerase II. Polyadenylation sequences can improve mRNA stability by adding a poly(A) tail to the 3' end of a coding sequence, thereby contributing to improved translation efficiency. In certain embodiments, BGHpA, hGHpA, or SV40pA may be utilized. In certain embodiments, a preferred embodiment of an expression construct includes a terminator element. Terminator elements can increase the amount of transcription and minimize transcription from the construct to another plasmid sequence by read-through.
[0073] In certain embodiments, the viral vector further comprises one or more insulator elements. The insulator elements may protect sequences expressed from the viral vector, such as effector elements and expressible elements, from integration site effects. Integration site effects occur through cis-acting elements in genomic DNA, meaning that the imported sequence is either expressed or not expressed (i.e., position effects; see, for example, Burgess-Beusse et al., PNAS., USA, 99:16433, 2002; and Zhan et al., Hum. Genet., 109:471, 2001). In certain embodiments, the viral import vector comprises one or more insulator elements in the 3'LTR, and upon provirus integration into the host genome, this insulator is integrated into both the 5'LTR and the 3'LTR during the replication of the 3'LTR. Insulators suitable for use in certain embodiments include the chicken β-globin insulator (see Chung et al., Cell 74:505, 1993; Chung et al., PNAS USA 94:575, 1997; and Bell et al., Cell 98:387, 1999), the SP10 insulator (Abhyankar et al., JBC 282:36143, 2007), or other small CTCF recognition sequences that function as enhancer-blocking insulators (Liu et al., Nature Biotechnology, 33:198, 2015).
[0074] In addition to the above, various types of suitable expression vectors are also known to those skilled in the art. These known expression vectors include commercially available expression vectors designed for general recombinant manipulation, such as plasmids that contain one or more reporter genes and the regulatory elements required for the reporter gene to be expressed in cells. Many vectors are commercially available, such as from Invitrogen, Stratagene, Clontech, etc., and are described in various accompanying guidebooks. In certain embodiments, suitable expression vectors include any plasmid, cosmid, or phage construct that can express the encoded gene in mammalian cells, such as the pUC plasmid system and the Bluescript plasmid system.
[0075] Particular embodiments of the vectors disclosed herein include those set forth in the table below. [Table 1]
[0076] Subcomponent sequences within a larger vector sequence can be readily identified by one of skill in the art based on the disclosure herein (see FIG. 3). The nucleotides between the identifiable subcomponents listed in the table above are restriction enzyme recognition sites used in construct assembly (cloning) and, in some cases, additional nucleotides with no identifiable function. These segments of the complete vector sequence can be adjusted using different cloning techniques and / or different vectors. Short palindromic sequences of six bases usually represent vector construction artifacts that are not critical to the function of the vector.
[0077] In certain embodiments, a vector (e.g., AAV) is selected that has a capsid that can cross the blood-brain barrier (BBB). In certain embodiments, the vector is engineered to contain a capsid that crosses the blood-brain barrier. Examples of AAVs with viral capsids that can cross the blood-brain barrier include AAV9 (Gombash et al., Front Mol Neurosci. 2014; 7:81), AAVrh.10 (Yang, et al., Mol Ther. 2014; 22(7): 1299-1309), AAV1R6, AAV1R7 (Albright et al., Mol Ther. 2018; 26(2): 510), rAAVrh.8 (Yang et al., supra), AAV-BR1 (Marchio et al., EMBO Mol Med. 2016; 8(6): 592), AAV-PHP.S (Chan et al., Nat Neurosci. 2017; 20(8): 1172), and AAV-PHP.B (Deverman et al., Nat Biotechnol. 2016; 34(2): 204), AAV-PPS (Chen et al., Nat Med. 2009; 15: 1215) and PHP.eB. In certain embodiments, the capsid of PHP.eB differs from that of AAV9 in that the amino acid residues from position 586 onwards, S-AQ-A (SEQ ID NO: 77), are changed to S-DGTLAVPFK-A (SEQ ID NO: 78), when comparing AAV9 as a reference. In certain embodiments, PHP.eb refers to the sequence of SEQ ID NO: 44.
[0078] AAV9 is a naturally occurring AAV serotype that, unlike many other naturally occurring serotypes, is able to cross the blood-brain barrier (BBB) upon intravenous injection. AAV9 transduces a wide area of the central nervous system (CNS), allowing for minimally invasive therapeutic approaches (Naso et al., BioDrugs. 2017; 31(4): 317). Such cases have been reported, for example, in connection with AveXis' clinical trial for the treatment of spinal muscular atrophy (SMA) syndrome (AVXS-101, NCT03505099) and the clinical trial for the treatment of CLN3-associated neuronal ceroid lipofuscinosis (NCT03770572).
[0079] AAVrh.10 is an AAV originally isolated from rhesus macaques that has weak human seroreactivity compared to other common serotypes used in gene delivery applications (Selot et al., Front Pharmacol. 2017; 8: 441) and has been evaluated in several clinical trials (LYS-SAF302, LYSOGENE and NCT03612869).
[0080] AAV1R6 and AAV1R7 are two variants isolated from a library of chimeric AAV vectors in which the capsid domain of AAVrh.10 is replaced by that of AAV1, which retain the ability to cross the BBB and transduce the central nervous system but show significantly reduced transduction of liver and vascular endothelium.
[0081] rAAVrh.8, also an AAV isolated from rhesus macaques, showed widespread transduction of glial and neuronal cells in clinically relevant areas following peripheral administration, with reduced peripheral tissue tropism compared to other vectors.
[0082] AAV-BR1 is an AAV2 variant that displays the NRGTEWD epitope (SEQ ID NO: 79) and was isolated during in vivo screening of a random AAV-display peptide library. AAV-BR1 exhibits high specificity with high transgene expression in the brain and minimal off-target affinity, including for the liver (Korbelin et al., EMBO Mol Med. 2016; 8(6): 609).
[0083] AAV-PHP.S (Addgene, Watertown, MA) is a variant of AAV9 generated by the CREATE method that encodes the 7-mer sequence QAVRTSL (sequence number 80) and transduces neurons of the enteric nervous system and potently transduces peripheral sensory afferents that project to the spinal cord and brainstem.
[0084] AAV-PHP.B (Addgene, Watertown, MA) is a variant of AAV9 generated by the CREATE method that encodes the 7-mer sequence TLAVPFK (SEQ ID NO: 81). AAV-PHP.B transfers genes throughout the CNS more efficiently than AAV9, transducing a large proportion of astrocytes and neurons in multiple CNS regions.
[0085] AAV-PPS is an AAV2 variant created by inserting the DSPAHPS epitope (SEQ ID NO: 82) into the capsid of AAV2, and exhibits dramatically improved brain tropism compared to AAV2.
[0086] For more information regarding capsids crossing the blood-brain barrier, see Chan et al., Nat. Neurosci. 2017 Aug: 20(8): 1172-1179.
[0087] (ii) Composition for Administration The artificial expression constructs and vectors (herein referred to as bioactive components) of the present disclosure can be formulated with carriers suitable for administration to cells, tissue slices, animals (e.g., mice and non-human primates), or humans. The bioactive components contained in the compositions described herein can be prepared in a neutral form, can be prepared as a free base, or can be prepared as a pharmacologically acceptable salt.
[0088] Pharmaceutically acceptable salts include the acid addition salts (formed from the free amino groups of the protein) which are formed with inorganic acids such as, for example, hydrochloric or phosphoric acid, or organic acids such as acetic, oxalic, tartaric, mandelic, etc. Also, salts formed with the free carboxyl groups can be derived from inorganic bases such as, for example, sodium, potassium, ammonium, calcium, or ferric hydroxides, or organic bases such as isopropylamine, trimethylamine, histidine, procaine, and the like.
[0089] Carriers for biologically active ingredients include solvents, dispersion media, vehicles, coating agents, diluents, isotonicity agents, absorption delaying agents, buffers, solutions, suspensions, colloids, etc. The use of such carriers for biologically active ingredients is well known in the art. Except insofar as a conventional media or agent is incompatible with the biologically active ingredients of the present invention, any conventional media or agent can be used in combination with the compositions described herein.
[0090] A "pharmacologically acceptable carrier" refers to a carrier that does not produce an allergic or similar untoward reaction when administered to a human, and in certain embodiments, when administered intravenously (e.g., into the retro-orbital plexus).
[0091] In certain embodiments, compositions of the invention can be formulated for intravenous, intraparenchymal, intraocular, intravitreal, parenteral, subcutaneous, intraventricular, intramuscular, intrathecal, intraspinal, intraperitoneal, oral or nasal inhalation, or for direct injection or administration into one or more cells, tissues or organs.
[0092] The compositions of the present invention may comprise liposomes, lipids, lipid complexes, microspheres, microparticles, nanospheres and / or nanoparticles.
[0093] The formation of liposomes and their use are widely known to those skilled in the art. Liposomes have been developed to improve serum stability and blood half-life (see, for example, U.S. Patent No. 5,741,516). In addition, various methods have been reported for using liposomes and liposome-like preparations as potential drug carriers (see, for example, U.S. Patent No. 5,567,434; U.S. Patent No. 5,552,157; U.S. Patent No. 5,565,213; U.S. Patent No. 5,738,868; and U.S. Patent No. 5,795,587).
[0094] The present disclosure also provides pharma- ceutically acceptable nanocapsule formulations of the bioactive ingredients of the present invention. In general, nanocapsule formulations can encapsulate compounds in a stable and reproducible manner (Quintanar-Guerrero et al., Drug Dev Ind Pharm 24(12):1113-1128, 1998; Quintanar-Guerrero et al., Pharm Res. 15(7):1056-1062, 1998; Quintanar-Guerrero et al., J. Microencapsul. 15(1):107-119, 1998; Douglas et al., Crit Rev Ther Drug Carrier Syst 3(3):233-261, 1987). To avoid side effects caused by large amounts of macromolecules being taken up into cells, such ultrafine particles can be designed using polymers that can be degraded in vivo. Biodegradable polyalkyl cyanoacrylate nanoparticles that meet such requirements are also envisioned for use in the present disclosure.Such microparticles can be easily prepared, and are described, for example, in Couvreur et al., J Pharm Sci 69(2):199-202, 1980; Couvreur et al., Crit Rev Ther Drug Carrier Syst. 5(1)1-20, 1988; zur Muhlen et al., Eur J Pharm Biopharm, 45(2):149-155, 1998; Zambaux et al., J Control Release 50(1-3):31-40, 1998; and US Patent No. 5,145,684.
[0095] Injectable compositions include sterile aqueous solutions or dispersions and sterile powders for extemporaneous preparation of sterile injectable solutions or dispersions (US Pat. No. 5,466,468). Injectable compositions delivered by injection are in the form of a sterile fluid to the extent that they can be delivered using a syringe. In certain embodiments, injectable compositions are usually stable during manufacturing and storage, and may contain one or more preservative compounds to prevent the contaminating action of microorganisms such as bacteria and fungi. The carrier may be a solvent or dispersion medium, which may include, for example, water, ethanol, polyol (e.g., glycerol, propylene glycol, liquid polyethylene glycol, and the like), and suitable mixtures thereof, and / or vegetable oils. To maintain proper fluidity, for example, a coating agent such as lecithin may be used, or in the case of dispersions, the particle size may be maintained to the required size, and / or a surfactant may be used. To prevent the action of microorganisms, various antibacterial and / or antifungal agents may be used, such as, for example, parabens, chlorobutanol, phenol, sorbic acid, thimerosal, and the like. In various embodiments, the injectable composition contains an isotonic agent, such as, for example, sugars or sodium chloride. Prolonged absorption of the injectable composition can be achieved by incorporating an agent that delays absorption, such as, for example, aluminum monostearate or gelatin, into the injectable composition. If necessary, an appropriate buffer may be added to the injectable composition, and the diluted liquid is first made isotonic with sufficient saline or glucose.
[0096] Dispersions may also be prepared in glycerol, liquid polyethylene glycols or mixtures thereof, or oils. As described herein, under ordinary conditions of storage and use, such preparations may contain a preservative to prevent the growth of microorganisms.
[0097] Sterile compositions can be prepared by mixing the physiologically active ingredient with any other ingredients (e.g., those mentioned above) in an appropriate amount of solvent and sterilizing by filtration. Dispersions are usually prepared by dispersing various sterilized physiologically active ingredients in a sterile solvent containing a basic dispersion medium and other necessary ingredients (e.g., those mentioned above). In the case of sterile powders for preparing sterile injectable solutions, a preferred method is to sterilize a solution containing the physiologically active ingredient and other desired ingredients in advance by filtration, and then vacuum drying or freeze-drying the solution to prepare a powder containing the physiologically active ingredient and other desired ingredients.
[0098] Oral compositions may be in liquid form, such as solutions, syrups, suspensions, etc., or may be provided as pharmaceutical products to be reconstituted with water or other suitable solvents before use. Such liquid preparations may be prepared by conventional methods using pharma- ceutical acceptable additives, such as suspending agents (e.g., sorbitol syrup, cellulose derivatives, or hydrogenated edible fats); emulsifying agents (e.g., lecithin or gum arabic); non-aqueous solvents (e.g., almond oil, ester oils, or fractionated vegetable oils); and preservatives (e.g., methyl p-hydroxybenzoate, propyl p-hydroxybenzoate, or sorbic acid). The composition of the present invention may be prepared, for example, in the form of tablets or capsules, by conventional methods using pharma- ceutically acceptable excipients, such as, for example, binders (e.g., pregelatinized corn starch, polyvinylpyrrolidone or hydroxypropylmethylcellulose); fillers (e.g., lactose, microcrystalline cellulose or calcium hydrogen phosphate); lubricants (e.g., magnesium stearate, talc or silica); disintegrants (e.g., potato starch or sodium starch glycolate); and wetting agents (e.g., sodium lauryl sulfate). Tablets may be coated by methods known in the art.
[0099] Compositions for inhalation can be delivered in the form of an aerosol spray from a pressurized pack or nebulizer using a suitable propellant, such as, for example, dichlorodifluoromethane, trichlorofluoromethane, dichlorotetrafluoroethane, carbon dioxide or other suitable gas. In the case of a pressurized aerosol, the dosage unit may be determined by providing a valve to deliver a metered amount. The compositions may be formulated into capsules or cartridges (e.g., gelatin capsules or cartridges) for use in an inhaler or nebulizer, which contain a powder mix of the compositions described herein and a suitable powder base, such as lactose or starch.
[0100] Additionally, compositions of the present invention include microchip devices (U.S. Pat. No. 5,797,898), ophthalmic formulations (Bourlais et al., Prog Retin Eye Res, 17(1):33-58, 1998), transdermal matrices (U.S. Pat. Nos. 5,770,219 and 5,783,208) and feedback controlled delivery (U.S. Pat. No. 5,697,899).
[0101] Supplementary active ingredients can also be included in the compositions of the present invention.
[0102] Typically, the compositions of the present invention may contain at least 0.1% or more of the physiologically active ingredient, but it goes without saying that the percentage of the physiologically active ingredient may vary and may conveniently range from 1% or 2% to 70% or 80% or more, or may range from 0.5 to 99%, based on the total weight or volume of the composition of the present invention. Of course, the amount of physiologically beneficial physiologically active ingredient in each composition may be adjusted so that a given unit dose of said compound provides an appropriate dosage. Factors such as solubility, bioavailability, biological half-life, route of administration, shelf life of the product, and other pharmacological considerations will be considered by those skilled in the art responsible for preparing such pharmaceutical formulations, and therefore various compositions and dosages may be desirable.
[0103] In certain embodiments, for human administration, compositions of the invention should meet sterility, pyrogenicity, general safety and purity standards as required by the U.S. Food and Drug Administration (FDA) or other relevant regulatory authorities in other countries.
[0104] (iii) Cell lines containing the artificial expression constructs The present disclosure includes cells comprising the artificial expression constructs described herein. Cells transformed with the artificial expression constructs can be used for a variety of purposes, such as neuroanatomical studies, evaluation of functional and / or non-functional proteins, drug screening to evaluate the regulatory properties of enhancers, etc.
[0105] While a variety of host cell lines can be used, in certain embodiments, the host cell is a mammalian cell. In certain embodiments, the artificial expression constructs are selected from the group consisting of eHGT_888m, eHGT_897m, 3xCore2-MGT_E51, eHGT_606h, MGT_E68, eHGT_889h, eHGT_1038m, eHGT_589m, eHGT_647m, eHGT_483m, 3xcore4_eHGT_606h, 3xcore5_eHGT_606h, 3xcore_eHGT_1039h, 3xCore_hTH, core4_eHGT_888m, 3Xcore4_eHGT_888m, 3X The enhancer and / or vector sequence is selected from core4_eHGT_888h and 3xcore1_eHGT_017m, and / or CN3406, CN3509, AiP1417, CN2436, AiP1240, CN3739, CN3889, CN2839, CN2847, CN2431, CN3058, CN3059, CN3829, CN3283, CN3863, CN4365, CN4830, and CN3551, and the host cell line is a human cell, a primate cell, or a mouse cell. Additionally, cell lines that can be used for gene transfer in the present disclosure include primary cell lines derived from living tissues such as rat or mouse brain, and organotypic cell cultures such as brain slices from animals, such as rat, mouse, non-human primate, or human neurosurgical tissues. The PC12 cell line (available from the American Type Culture Collection (ATCC), Manassas, VA) has been shown to express many neuronal marker proteins in response to nerve growth factor (NGF). The PC12 cell line is believed to be a neuronal cell line and is applicable for use in the present disclosure. JAR cells (available from the ATCC) are a platelet-derived cell line that express several neuronal genes, such as the serotonin transporter gene, and may be used in the embodiments described herein.
[0106] WO91 / 13150 describes various cell lines, including neuronal cell lines, and methods for their production. Similarly, WO97 / 39117 describes neuronal cell lines and methods for the production of such cell lines. The neuronal cell lines disclosed in these patent applications are applicable for use in the present disclosure.
[0107] In certain embodiments, the term "neuron" is used to describe any neuronal cell, related to neuronal cells, or including neuronal cells. Neuronal cells are defined by the characteristic of having an axon and dendrites. The term "neuron-specific" refers to something that is found in neuronal cells or cells derived therefrom, but is not found or is substantially absent in cells that are not derived from neuronal cells or in non-neuronal cells (e.g., glial cells such as astrocytes and oligodendrocytes); or activity that occurs in neuronal cells or cells derived therefrom, but is not found or is substantially absent in cells that are not derived from neuronal cells or in non-neuronal cells (e.g., glial cells such as astrocytes and oligodendrocytes).
[0108] In certain embodiments, non-neuronal cell lines such as mouse embryonic stem cells may be used. Cultured mouse embryonic stem cells can be transiently transfected with plasmid constructs to analyze the expression of gene constructs. Mouse embryonic stem cells are pluripotent undifferentiated cells. Mouse embryonic stem cells can be maintained in an undifferentiated state by leukemia inhibitory factor (LIF). Mouse embryonic stem cells can be induced to differentiate by removing LIF. Mouse embryonic stem cells form various types of differentiated cells in culture. Differentiation of mouse embryonic stem cells occurs through the expression of tissue-specific transcription factors, which allows the evaluation of the function of enhancer sequences (see, for example, Fiskerstrand et al., FEBS Lett 458: 171-174, 1999).
[0109] The method of differentiating stem cells into neuronal cells includes replacing the stem cell culture medium with a medium containing basic fibroblast growth factor (bFGF), heparin, N2 supplement (e.g., transferrin, insulin, progesterone, putrescine and selenite), laminin and polyornithine. The method of producing myelinating oligodendrocytes from stem cells is described in Hu, et al., 2009, Nat. Protoc. 4:1614-22. Bibel, et al., 2007, Nat. Protoc. 2:1034-43 describes a protocol for producing glutamatergic neurons from stem cells, and Chatzi, et al., 2009, Exp. Neurol. 217:407-16 describes a procedure for producing GABAergic neurons. This procedure includes exposing stem cells to all-trans retinoic acid for 3 days. GABAergic neurons, which account for 95% of the total cells, are then obtained by culturing in serum-free neuronal induction medium, such as neurobasal medium supplemented with B27, bFGF and EGF.
[0110] US Patent Publication No. 2012 / 0329714 describes the use of prolactin to increase neural stem cell numbers, and US Patent Publication No. 2012 / 0308530 describes a culture surface with amino groups that promotes neuronal differentiation into neurons, astrocytes, and oligodendrocytes. Thus, the fate of neural stem cells can be controlled by various extracellular factors. Commonly used extracellular factors include brain-derived growth factor (BDNF; Shetty and Turner, 1998, J. Neurobiol. 35:395-425); fibroblast growth factor (bFGF; U.S. Patent No. 5,766,948; FGF-1, FGF-2); neurotrophin 3 (nt-3) and neurotrophin 4 (nt-4) (Caldwell, et al., 2001, Nat. Biotechnol. 1;19:475-9); ciliary neurotrophic factor (CNTF); BMP-2 (U.S. Pat. Nos. 5,948,428 and 6,001,654); isobutyl-3-methylxanthine; leukemia growth inhibitory factor (LIF; U.S. Pat. No. 6,103,530); somatostatin; amphiregulin; neurotrophins (e.g., cyclic adenosine monophosphate); epidermal growth factor (EGF); dexamethasone (a glucocorticoid hormone); forskolin; ligands for the GDNF family of receptors; potassium; retinoic acid (U.S. Pat. No. 6,395,546); tetanus toxoid; and transforming growth factors alpha and TGF-beta (U.S. Pat. Nos. 5,851,832 and 5,753,506).
[0111] In certain embodiments, the yeast one-hybrid system may be used to identify compounds that inhibit specific protein-DNA interactions, such as, for example, the following transcription factors: eHGT_888m, eHGT_897m, 3xCore2-MGT_E51, eHGT_606h, MGT_E68, eHGT_889h, eHGT_1038m, eHGT_589m, eHGT_647m, eHGT_483m, 3xcore4_eHGT_606h, 3xcore5_eHGT_606h, 3xcore_eHGT_1039h, 3xCore_hTH, core4_eHGT_888m, 3Xcore4_eHGT_888m, 3Xcore4_eHGT_888h, or 3xcore1_eHGT_017m.
[0112] Transgenic animals are described below. Cell lines may be derived from such transgenic animals. For example, cell lines having artificial expression constructs integrated into their genomes can be obtained from primary tissue cultures derived from transgenic mice (e.g., as described below) (see, e.g., MacKenzie & Quinn, Proc Natl Acad Sci USA 96: 15251-15255, 1999).
[0113] (iv) Transgenic animals Another aspect of the disclosure includes a transgenic animal comprising in its genome an artificial expression construct comprising eHGT_888m, eHGT_897m, 3xCore2-MGT_E51, eHGT_606h, MGT_E68, eHGT_889h, eHGT_1038m, eHGT_589m, eHGT_647m, eHGT_483m, 3xcore4_eHGT_606h, 3xcore5_eHGT_606h, 3xcore_eHGT_1039h, 3xCore_hTH, core4_eHGT_888m, 3Xcore4_eHGT_888m, 3Xcore4_eHGT_888h and / or 3xcore1_eHGT_017m operably linked to a heterologous coding sequence. In certain embodiments, the genome of the transgenic animal comprises CN3406, CN3509, AiP1417, CN2436, AiP1240, CN3739, CN3889, CN2839, CN2847, CN2431, CN3058, CN3059, CN3829, CN3283, CN3863, CN4365, CN4830 and / or CN3551. In certain embodiments, when a non-integrating vector is utilized, the transgenic animal may be any of the following: eHGT_888m, eHGT_897m, 3xCore2-MGT_E51, eHGT_606h, MGT_E68, eHGT_889h, eHGT_1038m, eHGT_589m, eHGT_647m, eHGT_483m, 3xcore4_eHGT_606h, 3xcore5_eHGT_606h, 3xcore_eHGT_1039h, 3xCore_hTH, core4_eHGT_888m, The one or more cells comprise an artificial expression construct comprising 3Xcore4_eHGT_888m, 3Xcore4_eHGT_888h and / or 3xcore1_eHGT_017m, and / or CN3406, CN3509, AiP1417, CN2436, AiP1240, CN3739, CN3889, CN2839, CN2847, CN2431, CN3058, CN3059, CN3829, CN3283, CN3863, CN4365, CN4830 and / or CN3551.
[0114] A detailed description of the methods for producing transgenic animals is provided in U.S. Patent No. 4,736,866. The transgenic animals may be of any non-human species, but are preferably non-human primates (NHPs), sheep, horses, cows, pigs, goats, dogs, cats, rabbits, chickens; or rodents, such as guinea pigs, hamsters, gerbils, rats, mice, ferrets, etc.
[0115] In certain embodiments, by producing transgenic animals, organisms are obtained in which recombinant constructs are introduced into the same genome integration site of every cell.Therefore, the cell lines derived from such transgenic animals have consistent characteristics in that they have recombinant constructs in the same genome integration site of every cell, and therefore all of these cells undergo the same variegated position effect.In contrast, when gene is introduced into cell lines or primary cell cultures, heterologous expression of constructs is obtained.This method has the disadvantage that the expression of introduced DNA is affected by the specific genetic background of host animal.
[0116] As previously described in connection with cell lines, the artificial expression constructs of the present disclosure can be used to genetically modify mouse embryonic stem cells using techniques known in the art. Typically, the artificial expression constructs are introduced into cultured mouse embryonic stem cells. The transformed ES cells are then injected into blastocysts from a host mother, and the host embryo is reimplanted into the host mother. This procedure results in chimeric mice with tissues composed of cells derived from both embryonic stem cells present in the cultured cell line and embryonic stem cells present in the host embryo. Typically, mice are selected for isolating cultured ES cells used for gene transfer that have a different coat color than the host mouse into whose embryo the transformed cells are injected. Thus, the chimeric mice have a mixed coat color. If at least a portion of the germline tissue is derived from the genetically modified cells, the chimeric mice can then be crossed with an appropriate line to obtain offspring carrying the transgene.
[0117] In addition to the delivery methods described above, other methods of delivering artificial expression constructs to target cells or tissues or organs of animals, particularly cells, organs or tissues of mammalian vertebrates, are contemplated, including sonophoresis (e.g., ultrasound as described in U.S. Pat. No. 5,656,016); intraosseous injection (U.S. Pat. No. 5,779,708); microchip devices (U.S. Pat. No. 5,797,898); ophthalmic formulations (Bourlais et al., Prog Retin Eye Res, 17(1):33-58, 1998); transdermal matrices (U.S. Pat. Nos. 5,770,219 and 5,783,208); feedback controlled delivery (U.S. Pat. No. 5,697,899), as well as other delivery methods available and / or described elsewhere in this disclosure.
[0118] (v) How to use In certain embodiments, a composition comprising a bioactive ingredient described herein is administered to a subject to produce a physiological effect.
[0119] In certain embodiments, the present disclosure includes the use of the artificial expression constructs described herein to regulate the expression of a heterologous gene encoded in part or in its entirety downstream of an enhancer in a recombinant sequence. Accordingly, provided herein are methods of using the artificial expression constructs of the present disclosure in the research, study and future development of pharmaceuticals for the prevention, treatment or alleviation of symptoms of a disease, dysfunction or disorder.
[0120] Certain embodiments include a method of inducing expression of a gene in a dopaminergic neuron by administering to a subject an artificial expression construct, the artificial expression construct being any of the following described herein: eHGT_888m, eHGT_897m, 3xCore2-MGT_E51, eHGT_606h, MGT_E68, eHGT_889h, eHGT_1038m, eHGT_589m, eHGT_647m, eHGT_483m, 3xcore4_eHGT_606h, 3xcore5_eHGT_606h, 3xcore_eHGT_103 9h, 3xCore_hTH, core4_eHGT_888m, 3Xcore4_eHGT_888m, 3Xcore4_eHGT_888h and / or 3xcore1_eHGT_017m, and / or CN3406, CN3509, AiP1417, CN2436, AiP1240, CN3739, CN3889, CN2839, CN2847, CN2431, CN3058, CN3059, CN3829, CN3283, CN3863, CN4365, CN4830 and / or CN3551. The subject may be an isolated cell, a network of cells, a tissue section, an experimental animal, a veterinary animal or a human.
[0121] As is well known in the medical arts, the dose administered to a subject will depend on a variety of factors, such as the subject's body size, surface area and age, the particular compound administered, sex, duration and route of administration, general health, other drugs being administered concomitantly, etc. Although doses of the compounds of the present disclosure may vary, in certain embodiments, doses of the artificial expression constructs of the present disclosure may be administered in doses of 10 to 20 mg / kg. 5 ~10 100 In certain embodiments, patients receiving intravenous, intraparenchymal, intraspinal, retroorbital or intrathecal administration may receive 10 copies of the 6 ~10 22 A copy of the artificial expression construct can be injected.
[0122] "Effective amount" is the amount of a composition required to produce a desired physiological change in a subject. Effective amounts are often administered for research purposes. The effective amount disclosed herein is an amount that can produce a statistically significant effect in animal models, human studies, in vivo assays, or in vitro assays.
[0123] "Therapeutic treatment" includes treatment performed on a subject who shows symptoms or signs of disease, and is performed on a subject with the purpose of reducing or eliminating the signs or symptoms of disease. Therapeutic treatment can suppress, control or eliminate the presence or activity of disease or the cause of disease, and / or suppress, control or eliminate the side effects of disease. In certain embodiments, diseases that may be treated with aromatic L-amino acid decarboxylase (AADC) as a therapeutic payload for dopaminergic neurons include Parkinson's disease, multiple system atrophy with parkinsonism (MSA-P), and aromatic L-amino acid decarboxylase (AADC) deficiency. Other pathologies associated with decreased function of dopaminergic neurons (also referred to as "dopaminergic neuron-related diseases") include reverse resistant schizophrenia, Lesch-Nyhan syndrome (LNS), and attention deficit hyperactivity disorder (ADHD).
[0124] In certain embodiments, administration of a therapeutically effective amount of a composition of the present disclosure may be by a single administration, e.g., by a single injection of a sufficient number of expression constructs to provide a therapeutic benefit to the recipient subject. Alternatively, in some circumstances, it may be desirable to administer multiple or sequential doses of a composition of the present disclosure over a relatively short or long period of time.
[0125] In certain embodiments, the administration of therapeutically effective amount reduces the severity of Parkinson's disease.The reduction in severity is measured by measuring motor symptoms, evaluating the ability to perform daily functional activities, and symptomatic response to medication.The severity of Parkinson's disease is determined using several types of rating scales, including the Unified Parkinson's Disease Rating Scale (UPDRS), Hohn-Yahr severity classification, and Schwab-England Activities of Daily Living Scale.
[0126] The dose of the expression construct and the duration of administration of such compositions are determined by those skilled in the art who have the benefit of the teachings of the present invention. However, it is contemplated that administration of an effective amount of the composition of the present disclosure may be performed by a single administration, for example, by a single injection of a sufficient number of infectious particles to confer an effect on the subject. Alternatively, in some circumstances, it may be desirable to administer multiple or sequential administrations of the artificial expression construct composition or other genetic constructs over a relatively short or long period of time, and the decision to administer such may be determined by the person overseeing the administration of such compositions. For example, the number of infectious particles administered to a mammal may be as little as 10 or more times as necessary to achieve the intended effect. 7 pieces / ml, 10 8 pieces / ml, 10 9 pieces / ml, 10 10 pieces / ml, 10 11 pieces / ml, 10 12 pieces / ml, 10 13 This may be in the form of a single dose or two or more divided doses of cells / ml or more, and in certain embodiments, it may actually be desirable to administer two or more expression constructs in combination to achieve the desired effect.
[0127] In certain circumstances, it may be desirable to deliver the artificial expression construct in the form of an appropriately formulated composition as disclosed herein using a pipette or by retro-orbital injection, subcutaneous administration, intraocular administration, intravitreal administration, parenteral administration, subcutaneous administration, intravenous administration, intraparenchymal administration, intraventricular administration, intramuscular administration, intrathecal administration, intraspinal administration, intraperitoneal administration, oral administration, nasal inhalation, or direct administration or injection into one or more cells, tissues, or organs. Methods of administration may include those described in U.S. Patent No. 5,543,158; U.S. Patent No. 5,641,515, and U.S. Patent No. 5,399,363.
[0128] (vi) Kits and commercial packages The kits and commercial packages include an artificial expression construct as described herein. The artificial expression construct can be isolated. In certain embodiments, the components of the expression product can be separated from each other. In certain embodiments, the expression product is found within a vector, a viral vector, a cell, a tissue section or tissue sample, and / or a transgenic animal. Such kits may further include one or more reagents, restriction enzymes, peptides, therapeutic agents, pharmaceutical compounds, or a means for delivery of the compositions of the invention (e.g., a syringe, injection, etc.).
[0129] Embodiments of the kit or commercial package further include instructions for use of the components included in the kit or commercial package, e.g., in basic research, electrophysiological studies, neuroanatomical studies, and / or in the study and / or treatment of a disorder, disease or condition.
[0130] The following exemplary embodiments and experimental examples are provided to illustrate certain embodiments of the present disclosure. Those skilled in the art having reference to this disclosure will appreciate that various modifications can be made to the specific embodiments disclosed herein while still achieving the same or similar results without departing from the spirit and scope of the present disclosure.
[0131] (vii) Exemplary Embodiments 1. Artificial enhancers containing the cores of eHGT_606h, hTH, eHGT_1039h, eHGT_888m, eHGT_888h, MGT_E51 and / or eHGT_017m. 2. The artificial enhancer of embodiment 1, wherein the core comprises SEQ ID NO:2, SEQ ID NO:4, SEQ ID NO:6, SEQ ID NO:8, SEQ ID NO:83, SEQ ID NO:85, SEQ ID NO:90 or SEQ ID NO:88, or a sequence having at least 90% sequence identity with the sequence shown in SEQ ID NO:2, SEQ ID NO:4, SEQ ID NO:6, SEQ ID NO:8, SEQ ID NO:83, SEQ ID NO:85, SEQ ID NO:90 or SEQ ID NO:88. 3. An artificial enhancer according to embodiment 1 or 2, comprising 1, 2, 3, 4, 5, 6, 7, 8, 9 or 10 copies of MGT_E51, eHGT_606h, eHGT_1039h, hTH, eHGT_888m, eHGT_888h and / or eHGT_017m. 4. An artificial enhancer according to any one of embodiments 1 to 3, comprising 1, 2, 3, 4, 5, 6, 7, 8, 9 or 10 copies of a sequence as set forth in SEQ ID NO:2, SEQ ID NO:4, SEQ ID NO:6, SEQ ID NO:8, SEQ ID NO:83, SEQ ID NO:85, SEQ ID NO:90 and / or SEQ ID NO:88, or a sequence having at least 90% sequence identity with a sequence as set forth in SEQ ID NO:2, SEQ ID NO:4, SEQ ID NO:6, SEQ ID NO:8, SEQ ID NO:83, SEQ ID NO:85, SEQ ID NO:90 and / or SEQ ID NO:88. 5. An artificial enhancer described in any one of embodiments 1 to 4, comprising 1 copy, 2 copies, 3 copies, 4 copies, 5 copies, 6 copies, 7 copies, 8 copies, 9 copies or 10 copies of SEQ ID NO:2. 6. An artificial enhancer according to any one of embodiments 1 to 4, comprising 1 copy, 2 copies, 3 copies, 4 copies, 5 copies, 6 copies, 7 copies, 8 copies, 9 copies or 10 copies of SEQ ID NO:4. 7. An artificial enhancer according to any one of embodiments 1 to 4, comprising 1 copy, 2 copies, 3 copies, 4 copies, 5 copies, 6 copies, 7 copies, 8 copies, 9 copies or 10 copies of SEQ ID NO:6. 8. An artificial enhancer described in any one of embodiments 1 to 4, comprising 1 copy, 2 copies, 3 copies, 4 copies, 5 copies, 6 copies, 7 copies, 8 copies, 9 copies or 10 copies of SEQ ID NO:8. 9. An artificial enhancer according to any one of embodiments 1 to 4, comprising 1 copy, 2 copies, 3 copies, 4 copies, 5 copies, 6 copies, 7 copies, 8 copies, 9 copies or 10 copies of SEQ ID NO: 83. 10. An artificial enhancer according to any one of embodiments 1 to 4, comprising 1 copy, 2 copies, 3 copies, 4 copies, 5 copies, 6 copies, 7 copies, 8 copies, 9 copies or 10 copies of SEQ ID NO: 85. 11. An artificial enhancer according to any one of embodiments 1 to 4, comprising 1 copy, 2 copies, 3 copies, 4 copies, 5 copies, 6 copies, 7 copies, 8 copies, 9 copies or 10 copies of SEQ ID NO: 90. 12. An artificial enhancer described in any one of embodiments 1 to 4, comprising 1 copy, 2 copies, 3 copies, 4 copies, 5 copies, 6 copies, 7 copies, 8 copies, 9 copies or 10 copies of SEQ ID NO: 88. 13. An artificial enhancer according to embodiment 5, comprising three copies of sequence number 2. 14. An artificial enhancer according to embodiment 6, comprising three copies of sequence number 4. 15. An artificial enhancer described in embodiment 7, comprising three copies of SEQ ID NO:6. 16. An artificial enhancer described in embodiment 8, comprising three copies of SEQ ID NO:8. 17. An artificial enhancer according to embodiment 9, comprising three copies of SEQ ID NO:83. 18. An artificial enhancer according to embodiment 9, comprising one copy of SEQ ID NO:83. 19. An artificial enhancer described in embodiment 10, comprising three copies of SEQ ID NO:85. 20. An artificial enhancer described in embodiment 11, comprising three copies of SEQ ID NO:90. 21. An artificial enhancer described in embodiment 12, comprising three copies of SEQ ID NO:88. 22. An artificial enhancer described in embodiment 13, comprising the sequence shown in sequence number 3. 23. An artificial enhancer described in embodiment 14, comprising the sequence shown in sequence number 5. 24. An artificial enhancer described in embodiment 15, comprising the sequence shown in sequence number 7. 25. An artificial enhancer described in embodiment 16, comprising the sequence shown in sequence number 9. 26. An artificial enhancer according to embodiment 17, comprising the sequence shown in SEQ ID NO: 84. 27. An artificial enhancer described in embodiment 19, comprising the sequence shown in sequence number 86. 28. An artificial enhancer described in embodiment 20, comprising the sequence shown in sequence number 91. 29. An artificial enhancer described in embodiment 21, comprising the sequence shown in SEQ ID NO: 89. 30. An artificial expression construct comprising: (i) an enhancer selected from eHGT_888m, eHGT_897m, 3xCore2-MGT_E51, eHGT_606h, MGT_E68, eHGT_889h, eHGT_1038m, eHGT_589m, eHGT_647m, eHGT_483m, 3xcore4_eHGT_606h, 3xcore5_eHGT_606h, 3xcore_eHGT_1039h, 3xCore_hTH, core4_eHGT_888m, 3Xcore4_eHGT_888m, 3Xcore4_eHGT_888h and 3xcore1_eHGT_017m; (ii) a promoter (e.g., a minimal promoter); and (iii) heterologous coding sequence An artificial expression construct comprising: 31. The artificial expression construct of embodiment 30, wherein the heterologous coding sequence encodes an effector element or an expressible element. 32. The artificial expression construct of embodiment 30 or 31, wherein the effector element comprises a reporter protein or functional molecule. 33. The artificial expression construct of embodiment 32, wherein the reporter protein is a fluorescent protein. 34. The artificial expression construct of embodiment 32, wherein the functional molecule is a functional ion transporter, a functional enzyme, a functional transcription factor, a functional receptor, a functional membrane protein, a functional cell transport protein, a functional signaling molecule, a functional neurotransmitter, a functional calcium reporter, a functional channelrhodopsin, a functional CRISPR / Cas molecule, a functional editase, a functional guide RNA molecule, a functional microRNA, a functional homologous recombination donor cassette, or a functional designer receptor activated only by designer drugs (DREADD). 35. The artificial expression construct of embodiment 31, wherein the expressible element comprises a non-functional molecule. 36. The artificial expression construct of embodiment 35, wherein the non-functional molecule is a non-functional ion transporter, a non-functional enzyme, a non-functional transcription factor, a non-functional receptor, a non-functional membrane protein, a non-functional cell transport protein, a non-functional signaling molecule, a non-functional neurotransmitter, a non-functional calcium reporter, a non-functional channelrhodopsin, a non-functional CRISPR / Cas molecule, a non-functional editase, a non-functional guide RNA molecule, a non-functional microRNA, a non-functional homologous recombination donor cassette, or a non-functional designer receptor activated only by designer drugs (DREADD). 37. An artificial expression construct described in any one of embodiments 30 to 36, which is associated with a capsid that crosses the blood-brain barrier. 38. The artificial expression construct described in embodiment 37, wherein the capsid is PHP.eB, AAV-BR1, AAV-PHP.S, AAV-PHP.B or AAV-PPS. 39. An artificial expression construct according to any one of embodiments 30 to 38, comprising or encoding a skipping element. 40. The artificial expression construct described in embodiment 39, wherein the skipping element is a 2A peptide and / or an internal ribosome entry site (IRES). 41. The artificial expression construct of embodiment 40, wherein the 2A peptide is T2A, P2A, E2A or F2A. 42. eHGT_888m, eHGT_897m, 3xCore2-MGT_E51, eHGT_606h, MGT_E68, eHGT_889h, eHGT_1038m, eHGT_589m, eHGT_647m, eHGT_483m, 3xcore4_eHGT_606h, 3xcore5_eHGT_606h, 3xcore_eHGT_1039h, 3xCore_hTH, core4_eHGT_888m, 3Xcore4_eHGT_888m, 3Xcore4_eHGT_888h, 3xcore1_eHGT_017m, hsA2, AA 42. The artificial expression construct according to any one of embodiments 30 to 41, comprising or encoding a set of features selected from V, scAAV, rAAV, pAAV, minBglobin, CMV, minCMV, minRho, minRho*, fluorescent proteins (e.g., EGFP, SYFP2, GFP), Cre, iCre, dgCre, FlpO, tTA2, SP10 (e.g., 3xSP10), tag cassette, 10aa, nuclear transport protein, WPRE, WPRE3, hGHpA and / or BGHpA. 43. eHGT_888m-minBglobin-[heterologous coding sequence]-[post-transcriptional regulatory element]; eHGT_897m-minBglobin-[heterologous coding sequence]-[post-transcriptional regulatory element]; 3xCore2-MGT_E51-minBglobin-[heterologous coding sequence]-[post-transcriptional regulatory element]; eHGT_606h-minBglobin-[heterologous coding sequence]-[post-transcriptional regulatory element]; MGT_E68-minBglobin-[heterologous coding sequence]-[post-transcriptional regulatory element]; eHGT_889h-minBglobin-[heterologous coding sequence]-[post-transcriptional regulatory element]; eHGT_1038m-minBglobin-[heterologous coding sequence]-[post-transcriptional regulatory element]; eHGT_589m-minBglobin-[heterologous coding sequence]-[post-transcriptional regulatory element]; eHGT_647m-minBglobin-[heterologous coding sequence]-[post-transcriptional regulatory element]; eHGT_483m-minBglobin-[heterologous coding sequence]-[post-transcriptional regulatory element]; 3xcore4_eHGT_606h-minBglobin-[heterologous coding sequence]-[post-transcriptional regulatory element]; 3xcore5_eHGT_606h-minBglobin-[heterologous coding sequence]-[post-transcriptional regulatory element]; 3xcore_eHGT_1039h-minBglobin-[heterologous coding sequence]-[post-transcriptional regulatory element]; 3xCore_hTH-minBglobin-[heterologous coding sequence]-[post-transcriptional regulatory element]; core4_eHGT_888m-minBglobin-[heterologous coding sequence]-[post-transcriptional regulatory element]; 3Xcore4_eHGT_888m-minBglobin-[heterologous coding sequence]-[post-transcriptional regulatory element]; 3Xcore4_eHGT_888h-minBglobin-[heterologous coding sequence]-[post-transcriptional regulatory element]; 3xcore1_eHGT_017m-minBglobin-[heterologous coding sequence]-[post-transcriptional regulatory element]; eHGT_888m-minBglobin-[heterologous coding sequence]-WPRE3-BGHpA; eHGT_897m-minBglobin-[heterologous coding sequence]-WPRE3-BGHpA; 3xCore2-MGT_E51-minBglobin-[heterologous coding sequence]-WPRE3-BGHpA; eHGT_606h-minBglobin-[heterologous coding sequence]-WPRE3-BGHpA; MGT_E68-minBglobin-[heterologous coding sequence]-WPRE-hGHpA; eHGT_889h-minBglobin-[heterologous coding sequence]-WPRE3-BGHpA; eHGT_1038m-minBglobin-[heterologous coding sequence]-WPRE3-BGHpA; eHGT_589m-minBglobin-[heterologous coding sequence]-WPRE3-BGHpA; eHGT_647m-minBglobin-[heterologous coding sequence]-WPRE3-BGHpA; eHGT_483m-minBglobin-[heterologous coding sequence]-WPRE3-BGHpA; 3xcore4_eHGT_606h-minBglobin-[heterologous coding sequence]-WPRE3-BGHpA; 3xcore5_eHGT_606h-minBglobin-[heterologous coding sequence]-WPRE3-BGHpA; 3xcore_eHGT_1039h-minBglobin-[heterologous coding sequence]-WPRE3-BGHpA; 3xCore_hTH-minBglobin-[heterologous coding sequence]-WPRE3-BGHpA; core4_eHGT_888m-minBglobin-[heterologous coding sequence]-WPRE3-BGHpA; 3Xcore4_eHGT_888m-minBglobin-[heterologous coding sequence]-WPRE3-BGHpA; 3Xcore4_eHGT_888h-minBglobin-[heterologous coding sequence]-WPRE3-BGHpA; and 3xcore1_eHGT_017m-minBglobin-[heterologous coding sequence]-WPRE3-BGHpA 43. The artificial expression construct according to any one of embodiments 30 to 42, comprising or encoding a set of features selected from: 44. A vector comprising an artificial expression construct described in any one of embodiments 30 to 43. 45. The vector described in embodiment 44, which is a viral vector. 46. The vector described in embodiment 44 or 45, wherein the viral vector is a recombinant adeno-associated viral (AAV) vector. 47. An adeno-associated virus (AAV) vector comprising at least one heterologous coding sequence, wherein the heterologous coding sequence is under the transcriptional control of an enhancer and promoter selected from eHGT_888m, eHGT_897m, eHGT_606h, 3xCore2-MGT_E51, MGT_E68, eHGT_889h, eHGT_1038m, eHGT_589m, eHGT_647m, eHGT_483m, 3xcore4_eHGT_606h, 3xcore5_eHGT_606h, 3xcore_eHGT_1039h, 3xCore_hTH, core4_eHGT_888m, 3Xcore4_eHGT_888m, 3Xcore4_eHGT_888h and 3xcore1_eHGT_017m. 48. The AAV vector of embodiment 47, wherein the heterologous coding sequence encodes an effector element or an expressible element. 49. The AAV vector of embodiment 48, wherein the effector element comprises a reporter protein or functional molecule. 50. The AAV vector of embodiment 49, wherein the reporter protein is a fluorescent protein. 51. The AAV vector of embodiment 49, wherein the functional molecule is a functional ion transporter, a functional enzyme, a functional transcription factor, a functional receptor, a functional membrane protein, a functional cell transport protein, a functional signaling molecule, a functional neurotransmitter, a functional calcium reporter, a functional channelrhodopsin, a functional CRISPR / Cas molecule, a functional editase, a functional guide RNA molecule, a functional microRNA, a functional homologous recombination donor cassette, or a functional designer receptor activated only by designer drugs (DREADD). 52. The AAV vector described in embodiment 48, wherein the expressible element comprises a non-functional molecule. 53. The AAV vector of embodiment 52, wherein the non-functional molecule is a non-functional ion transporter, a non-functional enzyme, a non-functional transcription factor, a non-functional receptor, a non-functional membrane protein, a non-functional cell transport protein, a non-functional signaling molecule, a non-functional neurotransmitter, a non-functional calcium reporter, a non-functional channelrhodopsin, a non-functional CRISPR / Cas molecule, a non-functional editase, a non-functional guide RNA molecule, a non-functional microRNA, a non-functional homologous recombination donor cassette, or a non-functional designer receptor activated only by designer drugs (DREADD). 54. A transgenic cell comprising an expression construct or vector of any one of the preceding embodiments. 55. The transgenic cell of embodiment 54, which is a dopaminergic neuron. 56. The transgenic cell of embodiment 54 or 55, which is a mouse cell, a human cell or a non-human primate cell. 57. A non-human transgenic animal comprising an artificial expression construct, vector, or transgenic cell according to any one of the preceding embodiments. 58. A non-human transgenic animal described in embodiment 57, which is a mouse or a non-human primate. 59. An administrable composition comprising an expression construct, vector or transgenic cell according to any one of the preceding embodiments. 60. A kit comprising an artificial expression construct, vector, transgenic cell, transgenic animal and / or administrable composition according to any one of the preceding embodiments. 61. A method for expressing a gene in a target cell population in vivo or in vitro, comprising: 60. A method comprising the step of providing to a sample or subject comprising a target cell population an administrable composition of embodiment 59 at a sufficient dose and for a sufficient period of time, thereby expressing a gene in the target cell population. 62. The method of embodiment 61, wherein the gene encodes an effector element or an expressible element. 63. The method of embodiment 62, wherein the effector element comprises a reporter protein or functional molecule. 64. The method of embodiment 63, wherein the reporter protein is a fluorescent protein. 65. The method of embodiment 63, wherein the functional molecule is a functional ion transporter, a functional enzyme, a functional transcription factor, a functional receptor, a functional membrane protein, a functional cell transport protein, a functional signaling molecule, a functional neurotransmitter, a functional calcium reporter, a functional channelrhodopsin, a functional CRISPR / Cas molecule, a functional editase, a functional guide RNA molecule, a functional microRNA, a functional homologous recombination donor cassette, or a functional designer receptor activated only by designer drugs (DREADD). 66. The method of embodiment 62, wherein the expressible element comprises a non-functional molecule. 67. The method of embodiment 66, wherein the non-functional molecule is a non-functional ion transporter, a non-functional enzyme, a non-functional transcription factor, a non-functional receptor, a non-functional membrane protein, a non-functional cell transport protein, a non-functional signal transduction molecule, a non-functional neurotransmitter, a non-functional calcium reporter, a non-functional channelrhodopsin, a non-functional CRISPR / Cas molecule, a non-functional editase, a non-functional guide RNA molecule, a non-functional microRNA, a non-functional homologous recombination donor cassette, or a non-functional designer receptor activated only by designer drugs (DREADD). 68. The method of any one of embodiments 61 to 67, wherein the providing step comprises pipetting. 69. The method of embodiment 68, wherein the pipetting is performed on a brain slice. 70. The method of embodiment 69, wherein the brain slice is or is derived from the midbrain. 71. The method of embodiment 69 or 70, wherein the brain slice comprises dopaminergic neurons. 72. The method of any one of embodiments 69 to 71, wherein the brain slice is a mouse, human or non-human primate brain slice. 73. The method of any one of embodiments 68-72, wherein the providing step comprises administration to a living subject. 74. The method of embodiment 73, wherein the living subject is a human, a non-human primate, or a mouse. 75. The method of embodiment 73 or 74, wherein administration to the living subject is by injection. 76. The method of embodiment 75, wherein the injection is an intravenous injection, an intraparenchymal injection into brain tissue, an intracerebroventricular (ICV) injection, an intracisternal (ICM) injection or an intrathecal injection. 77. The method according to any one of embodiments 61 to 76, for use in the development of a treatment for a dopaminergic neuron-related disease or for the treatment of a dopaminergic neuron-related disease. 78. The method of embodiment 77, wherein the dopaminergic neuron-related disease comprises Parkinson's disease, multiple system atrophy with parkinsonism (MSA-P), aromatic L-amino acid decarboxylase (AADC) deficiency, adrenergic resistant schizophrenia, Lesch-Nyhan syndrome (LNS), and attention deficit hyperactivity disorder (ADHD). 79. The method of embodiment 77 or 78, wherein the gene encodes an aromatic L-amino acid decarboxylase (AADC). 80. The method of any one of embodiments 77 to 79, wherein the dopaminergic neuron-related disease comprises Parkinson's disease, multiple system atrophy with parkinsonism (MSA-P), or aromatic L-amino acid decarboxylase (AADC) deficiency. 81. The method of any one of embodiments 77 to 80, wherein the gene encodes aromatic L-amino acid decarboxylase (AADC) and the dopaminergic neuron-related disease comprises Parkinson's disease, multiple system atrophy with parkinsonism (MSA-P), or aromatic L-amino acid decarboxylase (AADC) deficiency. 82. The method of any one of embodiments 77-81, wherein the dopaminergic neuron-related disease comprises Parkinson's disease, and the treatment alleviates or eliminates symptoms of Parkinson's disease. 83. The method of embodiment 82, wherein the alleviation or elimination of symptoms of Parkinson's disease is measured by the Unified Parkinson's Disease Rating Scale (UPDRS), the Hoehn-Yahr classification, or the Schwab-England Activities of Daily Living Scale. 84. An artificial expression construct comprising CN3406, CN3509, AiP1417, CN2436, AiP1240, CN3739, CN3889, CN2839, CN2847, CN2431, CN3058, CN3059, CN3829, CN3283, CN3863, CN4365, CN4830 or CN3551.
[0132] (viii) Conclusion The nucleic acid and amino acid sequences provided herein are represented by the abbreviations used for nucleotide bases and amino acid residues as set forth in 37 CFR 1.831-1.835 and as set forth in WIPO Standard ST.26, effective July 1, 2022. Although only one strand is shown for each nucleic acid sequence, the complementary strand, if appropriate, is also included in the embodiments.
[0133] Variants of the sequences disclosed and referenced herein are also included herein. Guidelines for determining which amino acid residues can be substituted, inserted or deleted without losing biological activity can be determined using computer programs well known in the art, such as DNASTAR. TM Software (Madison, WI, USA) can be used to find the amino acid changes in the protein variants disclosed herein. The amino acid changes are preferably conservative amino acid changes, i.e., substitutions of similarly charged amino acids with each other or of uncharged amino acids with each other. Conservative amino acid changes include substitutions with members of a family of amino acids whose side chains are related.
[0134] Suitable conservative substitutions of amino acids in peptides or proteins are known to those skilled in the art, and such conservative substitutions can be made without generally altering the biological activity of the resulting molecule.Those skilled in the art will be familiar with the fact that generally, a single amino acid substitution in a non-essential region of a polypeptide will not substantially alter the biological activity (see, for example, Watson et al. Molecular Biology of the Gene, 4th Edition, 1987, The Benjamin / Cummings Pub. Co., p. 224). Naturally occurring amino acids are generally classified into conservative substitution families, specifically: Group 1: alanine (Ala), glycine (Gly), serine (Ser), and threonine (Thr); Group 2: (acidic): aspartic acid (Asp) and glutamic acid (Glu); Group 3: (acidic; also classified as polar, negatively charged residues and their amides): asparagine (Asn), glutamine (Gln), Asp, and Glu; Group 4: Gln and Asn; Group 5: (basic; also classified as polar, positively charged residues): arginine (Arg), lysine (Lys), and histidine (His); Group 6 (large aliphatic nonpolar residues): isoleucine (Ile), leucine (Leu), and ketone (K). Group 7 (polar uncharged): tyrosine (Tyr), Gly, Asn, Gln, Cys, Ser and Thr; Group 8 (large aromatic residues): phenylalanine (Phe), tryptophan (Trp) and Tyr; Group 9 (non-polar): proline (Pro), Ala, Val, Leu, Ile, Phe, Met and Trp; Group 11 (aliphatic): Gly, Ala, Val, Leu and Ile; Group 10 (small aliphatic residues that are non-polar or slightly polar): Ala, Ser, Thr, Pro and Gly; and Group 12 (sulfur-containing residues): Met and Cys. Further information can be found in Creighton (1984) Proteins, WH Freeman and Company.
[0135] In making such changes, the hydropathic index of amino acids may be taken into consideration. The importance of the hydropathic index of amino acids in conferring interactive biological function on a protein is widely understood in the art (Kyte and Doolittle, 1982, J. Mol. Biol. 157(1), 105-32). Each amino acid has been assigned a hydropathic index on the basis of its hydrophobicity and charge characteristics (Kyte and Doolittle, 1982). The hydrophobicity index of each amino acid is Ile (+4.5); Val (+4.2); Leu (+3.8); Phe (+2.8); Cys (+2.5); Met (+1.9); Ala (+1.8); Gly (-0.4); Thr (-0.7); Ser (-0.8); Trp (-0.9); Tyr (-1.3); Pro (-1.6); His (-3.2); glutamic acid (-3.5); Gln (-3.5); aspartic acid (-3.5); Asn (-3.5); Lys (-3.9); and Arg (-4.5).
[0136] It is well known in the art that substitution of a particular amino acid with another amino acid having a similar hydrophobicity index or hydrophobicity degree can also result in a protein with similar biological activity, i.e., a protein with biologically equivalent functionality. When making such changes, substitution of amino acids with hydrophobicity indices within ±2 is preferred, substitution of amino acids with hydrophobicity indices within ±1 is particularly preferred, and substitution of amino acids with hydrophobicity indices within ±0.5 is even more particularly preferred. Furthermore, it is well known in the art that substitution of similar amino acids can be effectively carried out based on hydrophilicity.
[0137] As detailed in U.S. Patent No. 4,554,101, each amino acid residue is assigned a hydrophilicity value, which is as follows: Arg (+3.0); Lys (+3.0); Aspartic acid (+3.0±1); Glutamic acid (+3.0±1); Ser (+0.3); Asn (+0.2); Gln (+0.2); Gly (0); Thr (-0.4); Pro (-0.5±1); Ala (-0.5); His (-0.5); Cys (-1.0); Met (-1.3); Val (-1.5); Leu (-1.8); Ile (-1.8); Tyr (-2.3); Phe (-2.5); Trp (-3.4). It is well known that certain amino acids can be substituted with other amino acids having a similar hydrophilicity value, and that such substitutions will result in biologically equivalent proteins, and in particular immunologically equivalent proteins. When making such changes, substitutions between amino acids whose hydrophilicity values are within the range of ±2 are preferred, substitutions between amino acids whose hydrophilicity values are within the range of ±1 are particularly preferred, and substitutions between amino acids whose hydrophilicity values are within the range of ±0.5 are even more particularly preferred.
[0138] As outlined above, amino acid substitutions may be made on the basis of the relative similarity of the amino acid side-chain substituents, for example, their hydrophobicity, hydrophilicity, charge, size, and the like.
[0139] As described elsewhere herein, variants of a gene sequence include codon-optimized variants, sequence polymorphisms, splice variants, and / or mutations that have no statistically significant effect on the function of the encoded product.
[0140] Variants of the protein, nucleic acid and gene sequences disclosed herein also include sequences having at least 70% sequence identity, at least 80% sequence identity, at least 85% sequence identity, at least 90% sequence identity, at least 95% sequence identity, at least 96% sequence identity, at least 97% sequence identity, at least 98% sequence identity or at least 99% sequence identity to the protein, nucleic acid or gene sequences disclosed herein.
[0141] "Percent sequence identity" refers to the relatedness of two or more sequences, as determined by comparing the sequences. In the art, "identity" also means the degree of relatedness between protein, nucleic acid or gene sequences, as determined by the matching between strings of protein, nucleic acid or gene sequences. "Identity" (often referred to as "similarity") can be readily calculated by known methods, including those described in Computational Molecular Biology (Lesk, AM, ed.) Oxford University Press, NY (1988); Biocomputing: Informatics and Genome Projects (Smith, DW, ed.) Academic Press, NY (1994); Computer Analysis of Sequence Data, Part I (Griffin, AM, and Griffin, HG, eds.) Humana Press, NJ (1994); Sequence Analysis in Molecular Biology (Von Heijne, G., ed.) Academic Press (1987); and Sequence Analysis Primer (Gribskov, M. and Devereux, J., eds.) Oxford University Press, NY (1992). Methods for determining identity are preferably designed to give the best match between the sequences tested. Methods for determining identity and similarity are codified in publicly available computer programs. Sequence alignment and identity calculations may be performed using the Megalign program (DNASTAR, Inc., Madison, Wis.) in the LASERGENE suite of bioinformatics computing software.Multiple alignment of sequences can also be performed using the Clustal format alignment method (Higgins and Sharp CABIOS, 5, 151-153 (1989) using default parameters (gap penalty=10, gap length penalty=10)). Related programs further include the GCG suite of programs (Wisconsin package version 9.0, Genetics Computer Group (GCG), Madison, Wisconsin); BLASTP, BLASTN, BLASTX (Altschul, et al., J. Mol. Biol. 215:403-410 (1990)); DNASTAR (DNASTAR, Inc., Madison, Wisconsin); and the FASTA program incorporating the Smith-Waterman algorithm (Pearson, Comput. Methods Genome Res., [Proc. Int. Symp.] (1994), Meeting Date 1992, 111-20. Editor(s): Suhai, Sandor. Publisher: Plenum, New York, NY). In this disclosure, when sequence analysis software is used for analysis, the analysis results are interpreted as being based on the "default values" that are the basis of the program. In this specification, "default values" refers to a set of numerical values or parameters that are preregistered in the software at the time of initialization of the software.
[0142] Variants also include nucleic acid molecules that hybridize to sequences disclosed herein under stringent hybridization conditions and have the same function as reference sequences.Exemplary stringent hybridization conditions include overnight incubation at 42°C in a solution containing 50% formamide, 5xSSC (750mM NaCl, 75mM trisodium citrate), 50mM sodium phosphate (pH 7.6), 5xDenhardt's solution, 10% dextran sulfate and 20μg / ml denatured salmon sperm DNA fragmented, followed by washing the filter with 0.1xSSC at 50°C.The stringency of hybridization and signal detection are mainly changed by adjusting the concentration of formamide (lower the percentage of formamide, lower the stringency), salt conditions or temperature. For example, moderately stringent conditions include overnight incubation at 37° C. in 6×SSPE (20×SSPE=3M NaCl; 0.2M NaH2PO4; 0.02M EDTA, pH 7.4), 0.5% SDS, 30% formamide, 100 μg / ml blocking salmon sperm DNA, followed by washing with 1×SSPE and 0.1% SDS at 50° C. Even lower stringency is achieved by performing stringent post-hybridization washes at high salt concentrations (e.g., 5×SSC). The above conditions can be varied by adding and / or substituting other blocking reagents used to reduce the background of hybridization experiments. Common blocking reagents include Denhardt's reagent, BLOTTO, heparin, denatured salmon sperm DNA, and commercially available proprietary preparations. The addition of certain blocking reagents may require some modification of the hybridization conditions described above due to compatibility issues.
[0143] The term "concatemerize" is used in a broad sense and means to link in a chain or to link in a series. The term is used to describe the linking of multiple nucleotide sequences to obtain a single nucleotide sequence or the linking of multiple amino acid sequences to obtain a single amino acid sequence. Also, "concatemerize" is understood to refer to "concatenation."
[0144] As will be appreciated by those of skill in the art, each embodiment disclosed herein comprises, consists essentially of, or consists of the particular components, steps, materials, or ingredients described. Thus, the terms "comprise" or "comprising" should be interpreted to mean "comprise, consist essentially of, or consist of." The transitional phrase "comprise" means, but is not limited to, the inclusion of any unrecited components, steps, materials, or ingredients, even if the amount is greater. The transitional phrase "consisting of" excludes any unrecited components, steps, materials, or ingredients. The transitional phrase "consisting essentially of" limits the scope of the embodiment to the recited components, steps, materials, or ingredients and those components, steps, materials, or ingredients that do not materially affect the embodiment. A significant effect is an effect that results in a statistically significant reduction in targeted expression in dopaminergic neurons using the artificial expression constructs disclosed herein.
[0145] In certain embodiments, "artificial" means not naturally occurring.
[0146] Unless otherwise indicated, all numerical values expressing quantities or properties of materials, such as molecular weight and reaction conditions, in the specification and claims are to be construed in all instances as modified by the term "about." Accordingly, unless otherwise indicated, the numerical parameters set forth in the specification and appended claims are approximations that may vary depending upon the desired properties sought to be obtained by the present invention. Without intending to limit the scope of the doctrine of equivalents to the scope of the claims, each numerical parameter should, at the very least, be construed in light of the number of reported significant digits and by applying ordinary rounding procedures. For clarity, the term "about," when used in conjunction with a stated value or range, has a meaning that would be reasonably interpreted by one of ordinary skill in the art, i.e., within ±20% of the stated value; within ±19% of the stated value; within ±18% of the stated value; within ±17% of the stated value; within ±16% of the stated value; within ±15% of the stated value; within ±14% of the stated value; within ±13% of the stated value; within ±12% of the stated value; within ±11% of the stated value; within ±10% of the stated value; within ±9% of the stated value; within ±8% of the stated value; within ±7% of the stated value; within ±6% of the stated value; within ±5% of the stated value; within ±4% of the stated value; within ±3% of the stated value; within ±2% of the stated value; or within ±1% of the stated value.
[0147] Notwithstanding that the numerical ranges and parameters setting forth the broad scope of the invention are approximations and approximate ranges, the numerical values set forth in the specific examples are reported as precisely as possible, however, all numerical values inherently contain certain errors necessarily resulting from the standard deviation associated with their respective testing measurements.
[0148] In the description of the present invention (particularly in the description of the claims below), the terms "a", "an", "the" and similar modifiers are intended to include both the singular and the plural unless otherwise indicated or the context clearly indicates otherwise. Numerical ranges described herein are intended to be a shorthand way of referring to each numerical value falling within the range individually. Unless otherwise indicated, each numerical value is described herein as if it were individually described herein. Any method described herein can be performed in any suitable order unless otherwise indicated or the context clearly indicates otherwise. The use of any examples or language of examples (e.g., "etc.") provided herein is intended to be for the purpose of illustrating the invention only and does not limit the scope of the invention as described in the claims. No term described herein should be construed as indicating any non-claimed element essential to the practice of the invention.
[0149] Groupings of other elements of the invention disclosed herein or of various embodiments of the invention should not be construed as limiting the invention. Members of each group may be described herein or in the claims individually or in combination with other members of the group or other elements described herein. It is anticipated that for reasons of convenience and / or patentability, one or more members of a group may be added to another group, or one or more members may be deleted from a group. When such additions or deletions are made, the specification includes groups that are constructed to satisfy the recitation of all Markush groups set forth in the appended claims.
[0150] Specific embodiments of the present invention are described herein, including those embodiments known to the inventors to be the best mode for carrying out the invention. Of course, those skilled in the art will readily appreciate that the embodiments described herein may be modified in various ways upon review of the above detailed description. The inventors anticipate that such modifications may be adopted by those skilled in the art, and intend that the present invention may be practiced in other ways than as specifically described herein. Accordingly, the present invention includes all modifications of the subject matter recited in the appended claims and all equivalents of the subject matter of the present invention to the extent permitted within the scope of applicable law. Moreover, the present invention includes all combinations of the above-described elements in any and all variations thereof, unless otherwise indicated or the context clearly dictates otherwise.
[0151] Additionally, throughout this specification, various patents, publications, journal articles and other documents are cited (references herein). Each reference cited herein is individually incorporated herein by reference for the teachings thereof as if it were a part of this specification.
[0152] Finally, the embodiments of the invention disclosed herein are to be considered as illustrative of the principles of the invention. Other modifications may be adopted within the scope of the invention. Thus, by way of example, but not by way of limitation, alternative configurations of the invention may be used in accordance with the teachings herein. Thus, the invention is not limited to what has been shown and described herein.
[0153] The details described herein are by way of example and are presented solely for the purpose of illustrating preferred embodiments of the present invention, to provide what is believed to be the most useful, and to facilitate an understanding of the principles and conceptual aspects of various embodiments of the present invention. In this regard, no structural details of the present invention are described in more detail than is necessary for a basic understanding of the present invention, and those skilled in the art will be able to easily understand how to actually embody some forms of the present invention by reading the description of the present invention in conjunction with the drawings and / or examples.
[0154] The definitions and explanations used in this disclosure are intended to control future interpretations, unless clearly and definitively changed in the following examples, or unless the interpretation is rendered meaningless or substantially meaningless by the meaning of the term. If the definition of a term is rendered meaningless or substantially meaningless by the interpretation of the term, please refer to the definition of the term from a dictionary known to those skilled in the art, such as Webster's Dictionary (3rd Edition) or Oxford Dictionary of Biochemistry and Molecular Biology (Ed. Anthony Smith, Oxford University Press, Oxford, 2004).
Claims
1. 1. An artificial expression construct comprising: (i) an enhancer comprising a sequence set forth in SEQ ID NO:10, SEQ ID NO:83, SEQ ID NO:84, SEQ ID NO:11, SEQ ID NO:12, SEQ ID NO:15, SEQ ID NO:16, SEQ ID NO:2, SEQ ID NO:3, SEQ ID NO:4, SEQ ID NO:5, SEQ ID NO:8, SEQ ID NO:9, SEQ ID NO:6, SEQ ID NO:7, SEQ ID NO:85, SEQ ID NO:86, SEQ ID NO:90, SEQ ID NO:91, SEQ ID NO:88 or SEQ ID NO:89, or a sequence having at least 95% sequence identity to a sequence set forth in SEQ ID NO:10, SEQ ID NO:83, SEQ ID NO:84, SEQ ID NO:11, SEQ ID NO:12, SEQ ID NO:15, SEQ ID NO:16, SEQ ID NO:2, SEQ ID NO:3, SEQ ID NO:4, SEQ ID NO:5, SEQ ID NO:8, SEQ ID NO:9, SEQ ID NO:6, SEQ ID NO:7, SEQ ID NO:85, SEQ ID NO:86, SEQ ID NO:90, SEQ ID NO:91, SEQ ID NO:88 or SEQ ID NO:89; (ii) a promoter; and (iii) Coding sequence An artificial expression construct comprising:
2. 2. The artificial expression construct of claim 1, wherein the coding sequence encodes a fluorescent protein or a neurotransmitter.
3. 10. The artificial expression construct of claim 1, wherein the construct is associated with a capsid that crosses the blood-brain barrier.
4. 4. The artificial expression construct of claim 3, wherein the capsid comprises PHP.eB, AAV9, AAVrh.10, AAV-BR1, AAV-PHP.S, AAV-PHP.B, or AAV-PPS.
5. 2. The artificial expression construct of claim 1, comprising or encoding a skipping element.
6. 6. The artificial expression construct of claim 5, wherein the skipping element comprises a T2A peptide, a P2A peptide, an E2A peptide, an F2A peptide, or an internal ribosome entry site (IRES).
7. 2. The artificial expression construct of claim 1, which is incorporated into a viral vector.
8. 8. The artificial expression construct of claim 7, wherein the viral vector comprises a recombinant adeno-associated viral (AAV) vector.
9. 1. A method for expressing a coding sequence in a target cell population in vitro, comprising: providing to a sample comprising a target cell population an administrable composition comprising an artificial expression construct in a sufficient dose and for a sufficient period of time, thereby expressing the coding sequence in said target cell population; The artificial expression construct comprises: (i) an enhancer comprising a sequence set forth in SEQ ID NO:10, SEQ ID NO:83, SEQ ID NO:84, SEQ ID NO:11, SEQ ID NO:12, SEQ ID NO:15, SEQ ID NO:16, SEQ ID NO:2, SEQ ID NO:3, SEQ ID NO:4, SEQ ID NO:5, SEQ ID NO:8, SEQ ID NO:9, SEQ ID NO:6, SEQ ID NO:7, SEQ ID NO:85, SEQ ID NO:86, SEQ ID NO:90, SEQ ID NO:91, SEQ ID NO:88 or SEQ ID NO:89, or a sequence having at least 95% sequence identity to a sequence set forth in SEQ ID NO:10, SEQ ID NO:83, SEQ ID NO:84, SEQ ID NO:11, SEQ ID NO:12, SEQ ID NO:15, SEQ ID NO:16, SEQ ID NO:2, SEQ ID NO:3, SEQ ID NO:4, SEQ ID NO:5, SEQ ID NO:8, SEQ ID NO:9, SEQ ID NO:6, SEQ ID NO:7, SEQ ID NO:85, SEQ ID NO:86, SEQ ID NO:90, SEQ ID NO:91, SEQ ID NO:88 or SEQ ID NO:89; (ii) a promoter; and (iii) Coding sequence A method comprising:
10. 1. A method for expressing a coding sequence in a target cell population in vivo, comprising: providing to a subject comprising a target cell population an administrable composition comprising an artificial expression construct in a sufficient dose and for a sufficient period of time, thereby expressing the coding sequence in said target cell population; The artificial expression construct comprises: (i) an enhancer comprising a sequence set forth in SEQ ID NO:10, SEQ ID NO:83, SEQ ID NO:84, SEQ ID NO:11, SEQ ID NO:12, SEQ ID NO:15, SEQ ID NO:16, SEQ ID NO:2, SEQ ID NO:3, SEQ ID NO:4, SEQ ID NO:5, SEQ ID NO:8, SEQ ID NO:9, SEQ ID NO:6, SEQ ID NO:7, SEQ ID NO:85, SEQ ID NO:86, SEQ ID NO:90, SEQ ID NO:91, SEQ ID NO:88 or SEQ ID NO:89, or a sequence having at least 95% sequence identity to a sequence set forth in SEQ ID NO:10, SEQ ID NO:83, SEQ ID NO:84, SEQ ID NO:11, SEQ ID NO:12, SEQ ID NO:15, SEQ ID NO:16, SEQ ID NO:2, SEQ ID NO:3, SEQ ID NO:4, SEQ ID NO:5, SEQ ID NO:8, SEQ ID NO:9, SEQ ID NO:6, SEQ ID NO:7, SEQ ID NO:85, SEQ ID NO:86, SEQ ID NO:90, SEQ ID NO:91, SEQ ID NO:88 or SEQ ID NO:89; (ii) a promoter; and (iii) Coding sequence Including, The method, wherein the subject is a non-human subject.
11. 11. The method of claim 9 or 10, wherein the coding sequence encodes a fluorescent protein or a neurotransmitter.
12. 10. The method of claim 9, wherein the providing step comprises pipetting performed on the brain slice.
13. 13. The method of claim 12, wherein the brain slice comprises dopaminergic neurons derived from the midbrain.
14. The method of claim 10 , wherein the providing step comprises administering to a living subject.
15. 15. The method of claim 14, wherein the living subject is a non-human primate or a mouse.
16. 15. The method of claim 14, wherein the administration is by injection.
17. 11. The method of claim 9 or 10 for use in the development of treatments for dopaminergic neuron-related diseases.
18. 18. The method of claim 17, wherein the dopaminergic neuron-related disease comprises Parkinson's disease, multiple system atrophy with parkinsonism (MSA-P), aromatic L-amino acid decarboxylase (AADC) deficiency, adverse-resistant schizophrenia, Lesch-Nyhan syndrome (LNS), and attention-deficit hyperactivity disorder (ADHD).
19. The method of claim 9 or 10, wherein the coding sequence encodes an aromatic L-amino acid decarboxylase (AADC).
20. An artificial enhancer comprising 1 copy, 2 copies, 3 copies, 4 copies, 5 copies, 6 copies, 7 copies, 8 copies, 9 copies or 10 copies of the sequence set forth in SEQ ID NO: 83, SEQ ID NO: 2, SEQ ID NO: 4, SEQ ID NO: 6, SEQ ID NO: 8, SEQ ID NO: 85, SEQ ID NO: 90 or SEQ ID NO: 88, or 1 copy, 2 copies, 3 copies, 4 copies, 5 copies, 6 copies, 7 copies, 8 copies, 9 copies or 10 copies of a sequence having at least 95% sequence identity to the sequence set forth in SEQ ID NO: 83, SEQ ID NO: 2, SEQ ID NO: 4, SEQ ID NO: 6, SEQ ID NO: 8, SEQ ID NO: 85, SEQ ID NO: 90 or SEQ ID NO:
88.
21. 21. The artificial enhancer of claim 20, comprising a sequence set forth in SEQ ID NO: 83, SEQ ID NO: 84, SEQ ID NO: 3, SEQ ID NO: 5, SEQ ID NO: 7, SEQ ID NO: 9, SEQ ID NO: 86, SEQ ID NO: 91 or SEQ ID NO: 89, or a sequence having at least 95% sequence identity with a sequence set forth in SEQ ID NO: 83, SEQ ID NO: 84, SEQ ID NO: 3, SEQ ID NO: 5, SEQ ID NO: 7, SEQ ID NO: 9, SEQ ID NO: 86, SEQ ID NO: 91 or SEQ ID NO: 89.