Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

34 results about "Dna storage" patented technology

DNA storage is the process of encoding and decoding binary data onto and from synthesized strands of DNA (deoxyribonucleic acid). In nature, DNA molecules contain genetic blueprints for living cells and organisms.

Preserving solution for stably preserving sample DNA (deoxyribonucleic acid) at normal temperature and preparation method of preserving solution

The invention discloses a preserving fluid for stably preserving sample DNA at normal temperature and a preparation method thereof, and belongs to the technical field of biological sample preservation. The preserving fluid comprises a lysis system, a nucleic acid protection system and a buffering and stabilizing system, the cracking system comprises a composite surfactant and an enzymolysis auxiliary agent; the composite surface active agent is prepared from polyether polyol fatty acid ester and cocamidopropyl hydroxy sulfobetaine; the enzymolysis auxiliary agent comprises lysozyme Lyso-V and protease K; the nucleic acid protection system comprises a nitrogen heterocyclic polyamine-carboxylic acid derivative and dextran sulfate; the nitrogen heterocyclic polyamine-carboxylic acid derivative comprises 1, 4, 7, 10-tetraazacyclododecane-N, N ', N' ', N ''tetraacetic acid, 1, 4, 7-triazacyclononane-N, N', N''-triacetic acid, disodium ethylene diamine tetraacetate-nitrogen heterocyclic derivative, and diethylenetriamine pentaacetic acid-piperazine derivative, and the nitrogen heterocyclic polyamine-carboxylic acid derivative comprises 1, 4, 7, 10-tetraazacyclododecane-N, N ', N '', N'' tetraacetic acid, 1, 4, 7-triazacyclononane-N, N', N ''-triacetic acid. The buffering and stabilizing system comprises an amphoteric buffering agent and a polymer stabilizer; the amphoteric buffering agent comprises 2-(N-morpholino) ethanesulfonic acid and N-tri (hydroxymethyl) methylglycine.
Owner:PEKING UNION MEDICAL COLLEGE HOSPITAL +1

DNA sequence reconstruction method and system based on multi-scale attention and contrast learning

The invention discloses a DNA sequence reconstruction method and system based on multi-scale attention and contrast learning, and relates to the technical field of DNA storage data reconstruction. Comprising the following steps: collecting a plurality of DNA sequence copies, screening out abnormal length sequences, and constructing a standardized clustering data set; performing one-hot coding and filling processing on the DNA sequence; extracting context dependent features and cross-sequence variation features; an Inter-Sequence multi-head attention mechanism is constructed to calculate the similarity between the sequences, and a weighted sequence tensor is generated; a global dependency relationship in the sequence is extracted through an Intra-Sequence multi-head attention mechanism; local offset features caused by insertion and deletion errors are extracted through a multi-size convolutional network; inputting a double-layer long-short-term memory network for sequence-level modeling, and outputting base reconstruction probability distribution; and constructing positive and negative sample pairs, calculating comparison loss, combining cross entropy loss to form a joint loss function, and outputting a high-precision DNA sequence reconstruction result. The method has high accuracy and robustness under the conditions of complex noise and multiple types of errors.
Owner:DALIAN UNIV

Invasive species identification method and device based on deep learning and DNA storage, and electronic equipment

The invention provides an invasive species identification method and device based on deep learning and DNA storage and electronic equipment, and relates to the technical field of invasive organism prevention and control, and the method comprises the steps: obtaining a to-be-identified invasive species image, and generating a first DNA sequence of a to-be-identified invasive species in the invasive species image through a deoxyribonucleic acid DNA encoder; obtaining a plurality of second DNA sequences to be hybridized from the invasive species database, and obtaining the hybridization yield of the first DNA sequence and each second DNA sequence by using a hybridization yield predictor; and determining the species category corresponding to the DNA sequence with the highest hybridization yield as the species category of the invasive species to be identified. According to the method provided by the invention, the strong feature extraction capability of deep learning is combined with the advantages of ultrahigh density and ultra-long stability of DNA as a data storage medium, so that end-to-end mapping and recognition from a species image to an exclusive DNA sequence thereof are realized.
Owner:BINZHOU MEDICAL COLLEGE

A DNA storage method and data information storage entity capable of realizing data random lossless reading

The application provides a DNA storage method and a data information storage entity capable of realizing random lossless reading of data, and the DNA storage method comprises the following steps: S1, preparation of a DNA double-stranded structure; S2, preparation of a data information storage entity; S3, random lossless reading of DNA data; and S4, recycling of the data information storage entity. According to the DNA storage method provided by the application, the random access function is realized by using the DNA double-stranded structure, and the long-term storage function is realized by using the outer wrapping hydrogel material, random reading in a file system with a large number of files is realized, and the original data is ensured not to be lost, so that the application has a wide application prospect in the field of DNA storage.
Owner:XIANGFU LAB

Design method of file system architecture oriented to DNA storage block equipment

The invention discloses a design method of a file system architecture oriented to DNA storage block equipment, and belongs to the technical field of DNA storage. The problem of providing more convenient, efficient and abundant DNA storage services for upper-layer users is solved. The method comprises the following steps: constructing a physical layer in a file system oriented to DNA storage block equipment, defining a primer block, and constructing a method for reading and writing DNA data input by DNA data input equipment oriented to the DNA storage block equipment on the basis of the primer block; constructing a file system layer in the file system oriented to the DNA storage block equipment, and establishing a method for converting the primer blocks into logic data blocks in the file system layer; and constructing the file system oriented to the DNA storage block device and a read-write delay optimization method of the DNA data input device, and completing the overall design of the file system oriented to the DNA storage block device. According to the invention, the data access efficiency of DNA storage is improved.
Owner:HARBIN INST OF TECH

DNA storage medium-oriented scalable vector graphics (SVG) image coding method and system

PendingCN121767470AImage codingError checkingAlgorithm
The invention provides a DNA storage medium-oriented scalable vector graphics (SVG) image coding method and system, and the method comprises the steps: reading an SVG file, analyzing the SVG image to obtain a document tree containing a plurality of nodes, and distributing structure information for representing the structure position of the nodes in the document tree for the nodes; coding the nodes to generate node fragments, wherein the node fragments comprise label codes generated according to labels of the nodes and attribute codes generated according to attributes of the nodes; aggregating the plurality of node fragments according to the label codes thereof to form at least one aggregation block, the aggregation block comprising header information and the plurality of node fragments, the header information being used for indexing the node fragments contained therein; the at least one aggregate block is converted into at least one DNA base sequence, and error correction information for error verification or repair is added to the at least one DNA base sequence. According to the method, the encoding length is reduced while the semantic integrity of the SVG is maintained, the error-resistant capability is improved, and the progressive decoding capability is provided.
Owner:SHANGHAI JIAOTONG UNIV

A biomimetic mineralized material with high stability for protecting DNA and a preparation method thereof

The application discloses a kind of biomimetic mineralization materials with high stable protection DNA and preparation method thereof.DNA molecule is self-assembled with divalent metal ion, and DNA / metal nanoparticle is obtained; polyelectrolyte layer is wrapped on the surface of DNA / metal nanoparticle by layer-by-layer assembly technology, and DNA / metal@LBL nanoparticle is obtained; polydopamine layer is deposited on the surface of DNA / metal@LBL nanoparticle, and DNA / Fe@LBL@PDA particle is obtained, which is biomimetic mineralization material with high stable protection DNA.The biomimetic mineralization material has the advantages of simple construction, high density and long-term stability, and can effectively protect the stability and integrity of DNA molecule in harsh external storage environment (such as ROS, high temperature, high salt, alkaline, humid, biological solvent and nuclease environment).The present application can provide technical support for DNA high-density storage and long-term stability storage, and has wide application value in the fields of biological medicine, genetic engineering, disease treatment and DNA storage.
Owner:FUZHOU UNIV

Robust multi-sequence reconstruction method based on maximum a posteriori probability in DNA storage

The application discloses a robust multi-sequence reconstruction method based on maximum posterior probability in DNA storage, and comprises the following steps: decoding each sequence in a cluster by using an improved BCJR decoder; converting each sequence in the cluster into corresponding time sequence representation; determining the weight of each sequence; and deducing a MAP algorithm formula of robust multi-sequence decoding, and integrating the sequence weight into joint posterior probability calculation to inhibit the influence of outlying sequences. By proposing a base IDS channel model, the BCJR decoding algorithm is improved, so that the defects that the existing statistical inference sequence reconstruction method is only suitable for binary error correction code are overcome. Therefore, the application can process the quaternary error correction of the base form in the DNA sequence.
Owner:TIANJIN UNIV

DNA (deoxyribonucleic acid) data storage molecular tag based on nanopore as well as preparation method and application of DNA data storage molecular tag

The invention relates to a DNA data storage molecular tag based on nanopores and a preparation method and application thereof, and belongs to the technical field of DNA storage. The molecular tag is prepared by modifying azido polysaccharide onto an alkyne-containing DNA chain through a click chemical reaction. Hybridizing and fixing the molecular tag and a DNA bracket chain containing a specific notch structure, and scanning by using a glass nanopore; when the molecular tag passes through the nanopore, a characteristic ionic current blocking signal is generated, and accordingly data coding is achieved. The molecular tag disclosed by the invention is low in cost and simple and convenient to operate; polysaccharide with wide sources and efficient click chemistry are utilized, so that the preparation cost and complexity are remarkably reduced; the physicochemical properties of polysaccharide molecules are stable, the consistency of read signal amplitudes is ensured, the decoding accuracy is high, the label structure stability is good, and a guarantee is provided for long-term storage of data; based on a reversible hybridization storage structure, data updating can be realized without synthesizing a new DNA chain.
Owner:CHONGQING UNIV OF POSTS & TELECOMM +1

DNA storage encoding method based on graph convolution network and self-attention mechanism

The application discloses a DNA storage coding method based on a graph convolution network and a self-attention mechanism, and belongs to the technical field of coding in DNA storage. Specifically, a DNA coding sequence meeting a combination constraint condition is predicted, first, existing DNA coding is screened and data is cleaned to construct a DNA storage coding training set; second, a prediction model based on a graph convolution neural network and a self-attention mechanism is trained, and the self-attention mechanism is used to capture the relationship of local DNA coding; then, the coding data processed into a graph is input into the prediction model to perform coding prediction meeting the combination constraint; finally, a DNA storage coding set meeting the condition is output. The application constructs a DNA storage coding training set, trains a graph convolution self-attention neural network, better captures the relationship between codings, and adopts a learning-based prediction model to perform DNA storage coding, so that the application has high coding efficiency when processing codings with complex constraints.
Owner:DALIAN UNIV OF TECH

Method and system for realizing DNA (Deoxyribonucleic Acid) storage by aiming at multi-rule rotation coding of Chinese text

The invention discloses a method and a system for realizing DNA storage by aiming at multi-rule rotation coding of Chinese texts, and relates to the technical field of DNA storage. Encoding the Chinese text by using a five-stroke font input method; mapping is carried out according to the positive and negative code tables, and an interval balance and dynamic detection strategy is introduced in the mapping process to control GC content distribution in intervals and the probability of occurrence of homopolymers between the intervals. Scattering and recombining the sequence by using block coding, and compressing by using RLE coding, wherein the RLE coding generates a sequence file and a run-length file; performing GC content constraint on the sequence file by using cross coding; and br compression is adopted for the run-length file to further improve the compression ratio. And generating a DNA sequence through rotary coding. According to the method, five-stroke coding, positive and negative code table mapping and sequence reconstruction strategies are introduced, so that the information storage density is higher, the local GC content of the DNA sequence is more stable, the homopolymer length is smaller, the unexpected motif proportion is lower, and the data has higher reliability and safety in the storage process.
Owner:DALIAN UNIV

Iterative decoding method and system of VT code without prior information in DNA storage and storage medium

The application discloses a kind of DNA storage in prior information VT code iterative decoding method, system and storage medium, it is related to DNA storage technical field.The method is first obtained systematic code word sequence by VT code, generates noisy reception sequence by ID model simulation channel transmission;Again, joint grid chart is constructed and initialized, and soft information and bit decision information are obtained by completing SISO decoding calculation, and channel parameters are updated based on MAP drift tracking;After convergence by iterative control, multiple independent estimated code words are obtained, and finally original information sequence is recovered by multi-sequence majority voting fusion.The application constructs decoding-estimation closed-loop iteration framework, does not need prior channel parameters, can adaptively learn channel characteristics, capture error hot spot, improve decoding robustness by combining multi-sequence fusion, realize original data high reliability recovery, and promote DNA storage practicality.
Owner:TIANJIN UNIV

A normal-temperature stable sample DNA storage solution and a preparation method thereof

This invention discloses a preservation solution for preserving sample DNA at room temperature and its preparation method, belonging to the field of biological sample preservation technology. The preservation solution comprises: a lysis system, a nucleic acid protection system, and a buffering and stabilizing system; the lysis system comprises a complex surfactant and an enzymatic hydrolysis aid; the complex surfactant comprises: polyether polyol fatty acid ester and cocamidopropyl hydroxysulfonate betaine; the enzymatic hydrolysis aid comprises lysozyme Lyso-V and proteinase K; the nucleic acid protection system comprises: nitrogen-heterocyclic polyamine-carboxylic acid derivatives and dextran sulfate; the nitrogen-heterocyclic polyamine-carboxylic acid derivatives comprise: 1,4,7,10-tetraazacyclododecane-N,N',N'',N'''tetraacetic acid, 1,4,7-triazacyclononane-N,N',N''-triacetic acid, disodium ethylenediaminetetraacetate-nitrocyclic derivative, and diethylenetriaminepentaacetic acid-piperazine derivative; the buffering and stabilizing system comprises: an amphoteric buffer and a polymeric stabilizer; the amphoteric buffer comprises: 2-(N-morpholino)ethanesulfonic acid and N-tris(hydroxymethyl)methylglycine.
Owner:PEKING UNION MEDICAL COLLEGE HOSPITAL +1

End-to-end DNA data storage coding method and system based on multi-constraint loss

The invention discloses an end-to-end DNA data storage coding method and system based on multiple constraint loss, and relates to the technical field of DNA storage. A complementary matrix convolution detection mechanism is designed in a coding stage, a hairpin structure is found through the complementary matrix convolution detection mechanism, a micronizable loss function is constructed, and a micronizable hairpin structure constraint loss function is realized. An interpretable error predictor based on Transform is trained through real sequencing data, a penalty loss function is constructed by means of an attention mechanism error-prone motif mask matrix to achieve error-prone motif suppression, and then GC content loss and homopolymer loss are fused to form a multi-constraint loss function. According to the method, an end-to-end architecture with a Transform auto-encoder as a main body is constructed, mapping of image data and a DNA sequence is realized through differentiable arithmetic coding, data reconstruction is completed in combination with noise injection and reverse decoding, and in a training stage, a deep Q network is adopted to carry out adaptive joint adjustment on weights of a multi-constraint loss function and mean square error reconstruction loss, so that data reconstruction is completed. And end-to-end optimization of model parameters is realized.
Owner:DALIAN UNIV

A method and system for converting ancient text sequences for DNA storage data based on five-stroke coding

The application discloses a kind of ancient text sequence conversion method and system for DNA storage data based on five-stroke encoding, comprising the following steps: ancient text is converted into five-stroke encoding alphabet sequence according to word;Five-stroke encoding alphabet sequence is compressed by Huffman ternary, and ternary digital string is obtained;Ternary digital string is converted into DNA base sequence by dynamic mapping rule;DNA base sequence is segmented, and positioning code is added to each segment, and the final DNA sequence set suitable for DNA synthesis and high-throughput sequencing is output;The final DNA sequence set is synthesized, sequenced, and the ancient text is restored by reverse operation, and its reliability is verified.The method can efficiently and accurately convert ancient text to DNA storage sequence, ensure biological compatibility and data integrity restoration, and meet the long-term DNA storage needs of ancient text.
Owner:NANJING UNIV OF SCI & TECH

Image DNA storage method for improving information density and biological stability

The invention provides an image DNA (deoxyribonucleic acid) storage method for improving information density and biological stability, which comprises the following steps of: partitioning an input image, performing two-dimensional discrete wavelet transform, performing quantization processing on a wavelet coefficient by adopting a self-adaptive quantization strategy, and converting the wavelet coefficient into a binary sequence to be coded; the method comprises the following steps: dividing a binary sequence to be coded into a byte according to every 8 bits, and carrying out dual-mode DNA coding by adopting an odd-even alternating mode and a segmentation structure of'feature segment-coding segment 'to generate a DNA information segment sequence; performing error correction on the coded DNA sequence based on preliminary correction of Hamming distance and RS code secondary error correction; marking the DNA sequence which still does not pass the CRC verification after multiple rounds of error correction as an unrepairable sequence, wherein the image block corresponding to the unrepairable sequence is an error block; and in combination with context information, the neighborhood features of the error blocks are learned by using a cross Transform context repair network, and the error blocks are repaired. According to the method, better balance is achieved in the aspects of image reconstruction quality, coding density and biological compatibility.
Owner:ZHENGZHOU UNIVERSITY OF LIGHT INDUSTRY

Methods and apparatus for DNA storage coding and decoding and rules thereof

The invention discloses a method and a device for DNA storage coding and decoding and rules thereof. The method comprises the following steps: performing single-molecule sequencing on a reference sequence to obtain actual sequencing data of single-molecule sequencing; comparing the actual sequencing data with reference data of the reference sequence, counting the frequency of sequencing errors of each sequence fragment with the length of k in the actual sequencing data, and calculating the proportion of the sequencing errors of each sequence fragment with the length of k in the actual sequencing data, namely the error rate; and taking the sequence fragments of which the error rates exceed a threshold value as limiting conditions to be eliminated. According to the method provided by the invention, the DNA storage coding and decoding steps are simplified, and the complexity of data processing is reduced through the time sequence of the threshold elimination step.
Owner:SHENZHEN HUADA GENE INST

A DNA storage readout method based on multiple hidden reference sequences

ActiveCN120596016BA-DNATesting Methods
The application discloses a DNA storage reading method based on multiple hidden reference sequences, which constructs multiple hidden reference sequences to realize reliable recovery of large fragment DNA from a read containing insertion / deletion errors; the constructed hidden reference sequence changes the reading of large fragment DNA storage from scratch into a low complexity resequencing problem; wherein, the watermark reference sequence quickly identifies low error reads, the skeleton reference sequence identifies reads containing insertion / deletion errors, and the decoding feedback reference sequence identifies reads for filling low coverage areas; the forward-backward algorithm of each read corrects the insertion / deletion errors in stages to realize reliable data recovery. The application realizes reliable data reading from reads containing insertion / deletion errors, and has the advantages that the constructed multiple hidden reference sequences identify reads with complementary characteristics of error rates in stages, the forward-backward algorithm of each read corrects the insertion / deletion errors in stages, and gradual reading is realized.
Owner:TIANJIN UNIV

High-reliability DNA storage coding and decoding method and related device

The invention discloses a high-reliability DNA storage coding and decoding method and a related device, and the method comprises the steps: obtaining a to-be-coded text information sequence, and filling a two-dimensional information matrix with the to-be-coded text information sequence; reading a character sequence in the two-dimensional information matrix from different direction dimensions, and generating a first DNA sequence and a second DNA sequence according to a preset direct mapping relation by using a reading result; sequentially segmenting the first DNA sequence and the second DNA sequence to obtain a plurality of short segments, endowing each short segment with a dimension label according to the direction dimension corresponding to each short segment, and endowing each short segment with a unique index according to the segmentation sequence of each short segment; according to the method and the related device, DNA storage coding and decoding can be carried out on Chinese information, and the technical problems that in the prior art, information coding and decoding efficiency is low, an error correction mechanism is complex, and cascade data are prone to being lost are solved.
Owner:SHENZHEN INST OF ADVANCED TECH CHINESE ACAD OF SCI

DNA data fingerprint generation method, device, equipment, medium and product

The invention discloses a DNA data fingerprint generation method, device and equipment, a medium and a product, and relates to the technical field of data fingerprint and biological information crossing. The method comprises the following steps: preprocessing a target DNA sequence file to extract a plurality of K-mer subsequences, filtering an extraction result to obtain a plurality of residual K-mer subsequences of which the coverage is higher than or equal to a coverage threshold, and selecting N feature subsequences from a filtering result, and finally, standardizing and organizing the selected sub-sequence and the coverage / and the hash value of the selected sub-sequence into a data fingerprint of the target DNA sequence file, so that efficient, compact and powerful data fingerprint capturing and data copying tracking traceability capabilities can be provided for a DNA information storage system, that is, the problems of data rapid identification and content retrieval and verification in the field of DNA information storage are solved; key support is provided for practicability of the DNA storage technology, and practical application and popularization are facilitated.
Owner:TIANJIN INST OF IND BIOTECH CHINESE ACADEMY OF SCI

A Method and System for Evaluating DNA Storage Sequencing Depth Based on Channel Simulation

This invention discloses a method and system for evaluating DNA storage sequencing depth based on channel simulation, addressing the problems of inaccurate predictions and limited guidance in existing uniform distribution models. The method includes: when sequencing data is available, fitting real data to obtain log-normal distribution parameters μ and σ; when sequencing data is unavailable, obtaining μ and σ through simulation modeling based on experimental parameters; and combining these parameters to calculate the decoding ratio of the coding strand in the noiseless channel and the sequencing depth boundary in the noisy channel. The system includes input, storage channel modeling, sequencing depth calculation, and output modules. The model of this invention closely reflects reality, provides accurate predictions, reduces DNA storage and retrieval costs, improves the success rate of single-sequencing decoding, and is easy to use, facilitating the practical application of the technology.
Owner:TIANJIN UNIV SYNTHETIC BIOLOGY FRONTIER RES INST

DNA storage random access address coding method and system

The invention relates to the technical field of DNA storage, and particularly discloses a DNA storage random access address coding method and system. A binary code is provided, so that a secondary structure of which the stem length is within a range of 3-9 is effectively avoided; codes and weak irrelevant codes obtained after irrelevant codes are improved are jointly used as component codes of a decoupling structure, the Knuth's balance technology is fused, and a random access address sequence meeting irrelevant requirements, secondary structure avoiding requirements and GC balance requirements at the same time is coded; the method comprises the following steps of: cascading p linear error correction code redundant bits for uncorrelated codes, and then jointly using the uncorrelated codes and codes with the same length as component codes of a decoupling structure, so as to encode an address sequence which is uncorrelated, avoids a secondary structure and has error correction capability at the same time. The coded high-quality address sequence has better performance on minimum free energy, melting temperature, GC content and average decoding success rate, and the robustness of a DNA storage system is improved.
Owner:DALIAN UNIV

A distributed array storage and microbe-based high capacity error correction DNA storage technology (Bio-RAID)

The present application relates to the technical field of biology, in particular to a distributed array storage and high-capacity error correction DNA storage method, and provides a method for DNA storage information, a method for information distributed storage by using microorganisms, a method for DNA storage verification, and a high-capacity error correction DNA storage array technology (Bio-RAID) based on a microorganism system. According to the present application, the information can be stored in vivo, and efficient replication and accurate restoration of the stored information can be realized. According to the present application, the stored information can be replicated in a low-cost and unlimited manner. According to the present application, after DNA is extracted from the living body, the stored information can be permanently stored.
Owner:INST OF MICROBIOLOGY CHINESE ACAD OF SCI

DNA movable type storage system and method

A DNA movable type storage system and method based on movable type printing including: (1) construct a physical DNA movable type library of data payloads and a physical DNA movable type library of indexes based on “DNA movable type codebook”, which consist of a variety of DNA oligonucleotides corresponding to all “DNA payload movable type elements” and “DNA index movable type elements” in the two libraries, respectively; (2) transcode from the storage binary data into corresponding “DNA movable type units”, each of which contains corresponding DNA payload movable type elements and related DNA index movable type elements; (3) link the abovementioned DNA payload and index movable type elements to form a physical “DNA movable type unit”, and put all generated DNA movable type units together for DNA storage, which cover all data information of the target file.
Owner:BEIJING INSTITUTE OF GENOMICS CHINESE ACADEMY OF SCIENCES (CHINA NATIONAL CENTER FOR BIOINFORMATION)

A DNA sequencing read clustering method and system based on representation learning

This invention belongs to the interdisciplinary field of DNA digital storage and bioinformatics, and discloses a method and system for DNA sequencing read clustering based on representation learning. The method includes: preprocessing the original DNA sequencing reads to obtain the sequencing reads themselves and corresponding variant 1 and variant 2 sets; performing representation learning on the DNA sequencing reads using a deep learning model based on the sequencing reads themselves and the corresponding variant 1 and variant 2 sets; and achieving DNA sequencing read clustering by fine-tuning the model based on the representation-learned DNA sequencing reads. This invention solves the problem of difficult clustering caused by sequencing errors in DNA storage.
Owner:GUANGZHOU UNIVERSITY

Encryption coding method and device in DNA storage, electronic equipment and storage medium

The application provides an encryption coding method and device in DNA storage, electronic equipment and storage medium, and relates to the technical field of DNA storage. The encryption coding method comprises the following steps: obtaining to-be-stored data, a first random number sequence and a second random number sequence generated by a hyperchaotic pseudo-random sequence generator; encrypting the to-be-stored data by using the first random number sequence to obtain source data; performing DNA Raptor coding on the source data, and introducing the second random number sequence into the DNA Raptor coding process as a random number seed to obtain at least one coded data; and converting each coded data into a corresponding DNA sequence based on a DNA base mapping rule, and generating a target DNA sequence from the DNA sequences corresponding to the coded data. The application solves the problem that data security is not considered in the DNA storage process in the related art.
Owner:SHENZHEN INST OF ADVANCED TECH CHINESE ACAD OF SCI

Quaternary dna storage method and system based on entropy coding and rotation constraint coding, storage medium and terminal

The application discloses a quaternary DNA storage method and system based on entropy coding and rotation constraint coding, a storage medium and a terminal. First, compared with traditional binary coding, quaternary coding can improve storage density, thereby storing more data in smaller space. Second, dynamic rotation coding ensures that the DNA sequence meets strict biochemical constraints, including homopolymer length limitation, GC content balance and avoidance of specific harmful sequences, optimizes the synthesis and sequencing process of DNA, and reduces the error rate. In addition, the use of tANS coding provides higher data compression efficiency, further reducing storage costs. These technical innovations not only improve the performance of large-scale data storage, but also enhance the security and accuracy of data decoding, ensuring the integrity of data after long-term storage.
Owner:TIANJIN UNIV

A decimal-based DNA storage encoding method, device and readable storage medium

The application discloses a kind of DNA storage encoding methods based on decimal, equipment and readable storage medium, belong to biological and information technology field.The encoding method of the present application selects 10 from 16 double-base code words in turn and encodes ten arabic numerals, binary data is converted into decimal number according to fixed length grouping, then the double-base code word is used to encode decimal number.When encoding data, first segment binary data, generate a certain proportion of the same length redundant sequence using RS error correction code, convert all sequences from decimal number to base, and add a certain length of base address index, sequence error correction code and primer sequence at both ends obtain the final base sequence, finally the obtained sequence is synthesized and saved by array chip method.The encoding method of the present application can more effectively correct the error of storage sequence and recover all the information stored, and the method breaks through the limitation of single set of DNA capacity, meets the demand of larger capacity storage.
Owner:任兆瑞

Physical unclonable anti-counterfeiting material based on DNA (Deoxyribose Nucleic Acid) nanostructure as well as preparation method and application thereof

The invention discloses a physical unclonable anti-counterfeiting material based on a DNA (Deoxyribose Nucleic Acid) nanostructure as well as a preparation method and application of the physical unclonable anti-counterfeiting material, and belongs to the technical field of information security and nano biology. According to the anti-counterfeiting material, three regular tetrahedron DNA monomers are connected through flexible double chains to form a DNA tripolymer, double-chain thermodynamic flexibility endows the anti-counterfeiting material with a nanoscale random geometric configuration, multi-dimensional encryption is achieved through different spectrum fluorophores modified randomly, and the information entropy reaches 4.3 bits / pixel. By combining deterministic self-assembly of the DNA nano-structure with dynamic flexibility, unification of deterministic preparation and random anti-counterfeiting characteristics is realized, and the method has excellent biocompatibility, high safety and strong expansibility, can be seamlessly compatible with a DNA computing system and a data storage system, and has wide application prospects. Multiple scenes such as information encryption authentication, biomedical authentication, flexible electronic anti-counterfeiting and DNA storage chip anti-counterfeiting are adapted, and an innovative platform is provided for the next generation of security authentication technology.
Owner:RENJI HOSPITAL AFFILIATED TO SHANGHAI JIAO TONG UNIV SCHOOL OF MEDICINE

CRISPR-Cas12a-based plasmid pool DNA database management method and system

The invention discloses a CRISPR-Cas12a-based plasmid pool DNA database management method and a CRISPR-Cas12a-based plasmid pool DNA database management system. The method comprises the following steps: S1, data writing: encoding data required to be stored by a user, assembling an encoded DNA sequence and a designed multi-layer address sequence into a plasmid map, and synthesizing and cloning to form a plasmid pool database; s2, data retrieval: S21, analyzing a target address; s22, assembling the one-way guide RNA and a Cas12a protein into a ribonucleoprotein complex; s23, carrying out specific cleavage on the plasmid pool database by utilizing a ribonucleoprotein complex; s24, enriching a cutting product; s25, performing amplification and sequencing on the enriched product, and restoring target data after decoding; s3, data deletion: S31, designing a one-way guide RNA (Ribonucleic Acid) aiming at a to-be-deleted data address sequence; s32, cutting a target plasmid through a Cas12a protein; and S33, separating and removing the cut plasmids through a gel electrophoresis method so as to realize data deletion, and promoting DNA storage from a static archiving mode to a dynamic database management mode.
Owner:TSINGHUA SHENZHEN INTERNATIONAL GRADUATE SCHOOL