Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

76 results about "Dna storage" patented technology

DNA storage is the process of encoding and decoding binary data onto and from synthesized strands of DNA (deoxyribonucleic acid). In nature, DNA molecules contain genetic blueprints for living cells and organisms.

Semantic intelligent enhanced DNA storage method for Internet of Things

The invention discloses a semantic intelligent enhanced DNA storage method for the Internet of Things, and belongs to the crossing field of artificial intelligence and biological information storage, the compression efficiency and robustness of DNA image storage are effectively improved through introduced semantic extraction and a multi-read screening mechanism of design, key information areas in images are extracted through semantic extraction, and the compression efficiency and robustness of DNA image storage are improved. The method has the advantages that redundant backgrounds are eliminated, coding lengths are reduced, high-quality sequences are selected from sequencing copies by the aid of scoring and sequence analysis through a multi-read screening mechanism, fault tolerance is enhanced, the method is applicable to efficient semantic storage in scenes of the internet of things, the problem of low storage efficiency due to excessive redundant information in original images in existing DNA image storage methods is solved, and the method is applicable to high-efficiency semantic storage in scenes of the internet of things. And a high bit error rate exists in the DNA synthesis and sequencing process, so that the image recovery quality is influenced.
Owner:YANGTZE DELTA REGION INST (QUZHOU) UNIV OF ELECTRONIC SCI & TECH OF CHINA

Information storage method based on DNA coding

The invention discloses an information storage method based on DNA (deoxyribonucleic acid) coding, which comprises the following steps: converting digital information to be stored into a quaternary code, converting the quaternary code into a DNA sequence consisting of adenine A, cytosine C, guanine G and thymine T according to a preset mapping rule, and expressing an additional information dimension by using a chemical modification state of a nucleotide pair; inserting error detection and correction codes in the DNA sequence every a predetermined number of basic group positions, segmenting the encoded DNA sequence into a plurality of fragments with the length of 100-150bp, and adding a specific recognition sequence and an index marker at two ends of each fragment; a reversible thermosensitive response DNA nanostructure is used as a carrier, DNA fragments are selectively combined to the carrier, and hierarchical storage and rapid retrieval of information are realized through a temperature gradient control system; the DNA storage information density can be remarkably improved, the service life of DNA storage can be remarkably prolonged, and a novel technical path is provided for large-scale, long-term and safe molecular information storage.
Owner:CHINA ELECTRONICS STANDARDIZATION INST

Preserving solution for stably preserving sample DNA (deoxyribonucleic acid) at normal temperature and preparation method of preserving solution

The invention discloses a preserving fluid for stably preserving sample DNA at normal temperature and a preparation method thereof, and belongs to the technical field of biological sample preservation. The preserving fluid comprises a lysis system, a nucleic acid protection system and a buffering and stabilizing system, the cracking system comprises a composite surfactant and an enzymolysis auxiliary agent; the composite surface active agent is prepared from polyether polyol fatty acid ester and cocamidopropyl hydroxy sulfobetaine; the enzymolysis auxiliary agent comprises lysozyme Lyso-V and protease K; the nucleic acid protection system comprises a nitrogen heterocyclic polyamine-carboxylic acid derivative and dextran sulfate; the nitrogen heterocyclic polyamine-carboxylic acid derivative comprises 1, 4, 7, 10-tetraazacyclododecane-N, N ', N' ', N ''tetraacetic acid, 1, 4, 7-triazacyclononane-N, N', N''-triacetic acid, disodium ethylene diamine tetraacetate-nitrogen heterocyclic derivative, and diethylenetriamine pentaacetic acid-piperazine derivative, and the nitrogen heterocyclic polyamine-carboxylic acid derivative comprises 1, 4, 7, 10-tetraazacyclododecane-N, N ', N '', N'' tetraacetic acid, 1, 4, 7-triazacyclononane-N, N', N ''-triacetic acid. The buffering and stabilizing system comprises an amphoteric buffering agent and a polymer stabilizer; the amphoteric buffering agent comprises 2-(N-morpholino) ethanesulfonic acid and N-tri (hydroxymethyl) methylglycine.
Owner:PEKING UNION MEDICAL COLLEGE HOSPITAL +1

DNA sequence reconstruction method and system based on multi-scale attention and contrast learning

The invention discloses a DNA sequence reconstruction method and system based on multi-scale attention and contrast learning, and relates to the technical field of DNA storage data reconstruction. Comprising the following steps: collecting a plurality of DNA sequence copies, screening out abnormal length sequences, and constructing a standardized clustering data set; performing one-hot coding and filling processing on the DNA sequence; extracting context dependent features and cross-sequence variation features; an Inter-Sequence multi-head attention mechanism is constructed to calculate the similarity between the sequences, and a weighted sequence tensor is generated; a global dependency relationship in the sequence is extracted through an Intra-Sequence multi-head attention mechanism; local offset features caused by insertion and deletion errors are extracted through a multi-size convolutional network; inputting a double-layer long-short-term memory network for sequence-level modeling, and outputting base reconstruction probability distribution; and constructing positive and negative sample pairs, calculating comparison loss, combining cross entropy loss to form a joint loss function, and outputting a high-precision DNA sequence reconstruction result. The method has high accuracy and robustness under the conditions of complex noise and multiple types of errors.
Owner:DALIAN UNIV

Invasive species identification method and device based on deep learning and DNA storage, and electronic equipment

The invention provides an invasive species identification method and device based on deep learning and DNA storage and electronic equipment, and relates to the technical field of invasive organism prevention and control, and the method comprises the steps: obtaining a to-be-identified invasive species image, and generating a first DNA sequence of a to-be-identified invasive species in the invasive species image through a deoxyribonucleic acid DNA encoder; obtaining a plurality of second DNA sequences to be hybridized from the invasive species database, and obtaining the hybridization yield of the first DNA sequence and each second DNA sequence by using a hybridization yield predictor; and determining the species category corresponding to the DNA sequence with the highest hybridization yield as the species category of the invasive species to be identified. According to the method provided by the invention, the strong feature extraction capability of deep learning is combined with the advantages of ultrahigh density and ultra-long stability of DNA as a data storage medium, so that end-to-end mapping and recognition from a species image to an exclusive DNA sequence thereof are realized.
Owner:BINZHOU MEDICAL COLLEGE

Error correction systems and methods for DNA storage

ActiveUS12373283B2Sequence analysisRedundant data error correctionAlgorithmDocument representation
A DNA-based storage system includes an error correction system operable to: (a) identify a DNA codeword from a DNA sequencing operation; (b) calculate an initial syndrome weight; (c) determine that the initial syndrome weight is greater than a predetermined threshold; (d) perform an alignment alteration in the information segment by: (i) selecting a skew point within the information segment; (ii) performing an indel operation on the information segment at the skew point; (iii) calculating a modified syndrome weight; (iv) comparing the initial syndrome weight with the modified syndrome weight; and (v) incorporating the indel operation into the information segment when the comparing indicates an improvement in the modified syndrome weight; (e) decode the modified codeword; and (f) transmit the contents of the output file to a computing device, the output file representing user data stored within the DNA molecule.
Owner:WESTERN DIGITAL TECHNOLOGIES INC

Multi-channel extensible automatic sample loading system oriented to DNA storage and calculation and control method

The invention discloses a multi-channel extensible automatic sample loading system for DNA storage and calculation and a control method. The system comprises a gas path control device, a multi-channel sample introduction device, a liquid flow control device, a multi-liquid-path liquid collection device, a micro-fluidic chip and an upper computer, operation control over the whole system is provided through the upper computer, a user can adjust the sample injection sequence and the reaction time of each channel, the liquid flow and the gas flow of each channel are monitored and adjusted in real time in an experiment, and multi-channel and high-precision automatic sample injection is achieved. According to the invention, accurate injection and automatic control of samples are realized, the efficiency and precision of sample treatment are greatly improved, and high automation, high-precision sample injection and multi-channel low cross contamination control capability are realized; the expandability is high, and modular expansion to hundreds of channels or even hundreds of channels is supported; meanwhile, the method is compatible with various application scenes, and is suitable for the fields of high-throughput DNA calculation, DNA information storage, biomedical detection and the like.
Owner:SHANGHAI JIAOTONG UNIV

A DNA-encoding-based information storage method

The present invention discloses an information storage method based on DNA coding. The method comprises the following steps: converting digital information to be stored into a quaternary code, and converting the quaternary code into a DNA sequence consisting of adenine A, cytosine C, guanine G, and thymine T according to a preset mapping rule, and using the chemical modification state of the nucleotide pairs to represent an additional information dimension; inserting error detection and correction codes at every predetermined number of base positions in the DNA sequence, dividing the encoded DNA sequence into multiple fragments of 100-150 bp in length, and adding specific recognition sequences and index markers at both ends of each fragment; utilizing a reversibly thermoresponsive DNA nanostructure as a carrier, selectively binding the DNA fragments to the carrier, and realizing hierarchical storage and rapid retrieval of information through a temperature gradient control system; the present invention can significantly improve the information density and lifespan of DNA storage, and provides a new technical path for large-scale, long-term, and secure molecular information storage.
Owner:CHINA ELECTRONICS STANDARDIZATION INST

Method for prolonging data storage time and application

PendingCN121343982ADNA preparationDNA stabilityEngineering
The invention relates to the technical field of DNA information storage, and discloses a method for prolonging data storage time and application, the method comprises the following steps: data is converted into a DNA sequence with a preset length, DNA is synthesized and stored, primary amino groups in basic groups of the DNA are independently modified by protective groups, the protective groups are selected from-COR, and the primary amino groups in the basic groups of the DNA are independently modified by the protective groups. R is selected from alkyl of C1-C6, substituted or unsubstituted naphthenic base of C3-C8 or substituted or unsubstituted phenyl. According to the method for prolonging the data storage time, the degradation rate of DNA can be remarkably inhibited, the stability of the DNA can be improved so as to realize long-time stable storage of the data, and the method is particularly suitable for forming a DNA storage scene with limited conditions, namely, the method is suitable for DNA synthesis in large companies with advanced technologies and equipment and also suitable for large companies with advanced technologies and equipment. The method is also suitable for DNA synthesis of small companies, and has application prospects and potential.
Owner:SHANGHAI DYNASTYGENE CO

An Optimization Method for DNA Storage Encoding Based on the RBS Algorithm

The present invention discloses a method for optimizing DNA storage coding based on the RBS algorithm. This method first converts binary data into a base sequence and divides the sequence into three files according to base frequencies. Secondly, three methods, namely data block merging, RH hybrid coding, and character merging, are respectively used to encode the three files. Finally, rotation coding is used to convert binary data into a base sequence and divide it into short sequences of 120 nt each. The codewords generated by the present invention have a higher information storage density and better coding quality in DNA storage, improving the efficiency, practicality, and stability of DNA storage.
Owner:DALIAN UNIV

A DNA information storage method with preview function

The application discloses a DNA information storage method with a preview function, comprising the following steps: S1, generating a binary sequence of an original file to be stored and a corresponding binary sequence of a preview file; S2, converting the binary sequence of the original file and the corresponding binary sequence of the preview file into a DNA base sequence of the original file and a DNA base sequence of the corresponding preview file, so that the whole DNA base sequence meets a constraint of synthetic biology; S3, judging whether a first type of DNA information sequence construction method or a second type of DNA information sequence construction method is adopted; S4, executing the first type of DNA information sequence construction method; and S5, executing the second type of DNA information sequence construction method. Compared with the prior art, the application realizes the file preview function on a traditional computer in a DNA information storage system, and brings great convenience to a file retrieval process of DNA storage.
Owner:TIANJIN UNIV

A DNA storage method and data information storage entity capable of realizing data random lossless reading

The application provides a DNA storage method and a data information storage entity capable of realizing random lossless reading of data, and the DNA storage method comprises the following steps: S1, preparation of a DNA double-stranded structure; S2, preparation of a data information storage entity; S3, random lossless reading of DNA data; and S4, recycling of the data information storage entity. According to the DNA storage method provided by the application, the random access function is realized by using the DNA double-stranded structure, and the long-term storage function is realized by using the outer wrapping hydrogel material, random reading in a file system with a large number of files is realized, and the original data is ensured not to be lost, so that the application has a wide application prospect in the field of DNA storage.
Owner:XIANGFU LAB

DNA sequence assembly method and system based on dynamic variable-order unitg-level graph

The invention discloses a DNA sequence assembly method and system based on a dynamic variable order unitg-level graph, and relates to the technical field of bioinformatics and DNA storage. According to the method, a pseudo genome and a source perception k-mer index are constructed, a Mid-Max and Min-Mid two-stage variable order expansion strategy is adopted, a k value is dynamically adjusted to enhance the graph structure connectivity, and the problems that under the condition of low coverage rate or high error rate, an existing de Bruijn graph method is prone to breakage and path fuzziness is prone to being generated are effectively solved. According to the method, a hidden path is accurately repaired through node connection and splitting operation, and redundancy k-mer is filtered in combination with index continuity and prefix similarity, so that the assembly integrity and accuracy are improved, and meanwhile, the robustness of DNA data reconstruction is remarkably enhanced; the method is suitable for various high-noise and low-coverage-rate scenes such as genome assembly and DNA storage and reconstruction, and particularly shows excellent performance in practical application with high requirements on data integrity and reliability.
Owner:DALIAN UNIV

Design method of file system architecture oriented to DNA storage block equipment

The invention discloses a design method of a file system architecture oriented to DNA storage block equipment, and belongs to the technical field of DNA storage. The problem of providing more convenient, efficient and abundant DNA storage services for upper-layer users is solved. The method comprises the following steps: constructing a physical layer in a file system oriented to DNA storage block equipment, defining a primer block, and constructing a method for reading and writing DNA data input by DNA data input equipment oriented to the DNA storage block equipment on the basis of the primer block; constructing a file system layer in the file system oriented to the DNA storage block equipment, and establishing a method for converting the primer blocks into logic data blocks in the file system layer; and constructing the file system oriented to the DNA storage block device and a read-write delay optimization method of the DNA data input device, and completing the overall design of the file system oriented to the DNA storage block device. According to the invention, the data access efficiency of DNA storage is improved.
Owner:HARBIN INST OF TECH

DNA storage medium-oriented scalable vector graphics (SVG) image coding method and system

The invention provides a DNA storage medium-oriented scalable vector graphics (SVG) image coding method and system, and the method comprises the steps: reading an SVG file, analyzing the SVG image to obtain a document tree containing a plurality of nodes, and distributing structure information for representing the structure position of the nodes in the document tree for the nodes; coding the nodes to generate node fragments, wherein the node fragments comprise label codes generated according to labels of the nodes and attribute codes generated according to attributes of the nodes; aggregating the plurality of node fragments according to the label codes thereof to form at least one aggregation block, the aggregation block comprising header information and the plurality of node fragments, the header information being used for indexing the node fragments contained therein; the at least one aggregate block is converted into at least one DNA base sequence, and error correction information for error verification or repair is added to the at least one DNA base sequence. According to the method, the encoding length is reduced while the semantic integrity of the SVG is maintained, the error-resistant capability is improved, and the progressive decoding capability is provided.
Owner:SHANGHAI JIAOTONG UNIV

A biomimetic mineralized material with high stability for protecting DNA and a preparation method thereof

The application discloses a kind of biomimetic mineralization materials with high stable protection DNA and preparation method thereof.DNA molecule is self-assembled with divalent metal ion, and DNA / metal nanoparticle is obtained; polyelectrolyte layer is wrapped on the surface of DNA / metal nanoparticle by layer-by-layer assembly technology, and DNA / metal@LBL nanoparticle is obtained; polydopamine layer is deposited on the surface of DNA / metal@LBL nanoparticle, and DNA / Fe@LBL@PDA particle is obtained, which is biomimetic mineralization material with high stable protection DNA.The biomimetic mineralization material has the advantages of simple construction, high density and long-term stability, and can effectively protect the stability and integrity of DNA molecule in harsh external storage environment (such as ROS, high temperature, high salt, alkaline, humid, biological solvent and nuclease environment).The present application can provide technical support for DNA high-density storage and long-term stability storage, and has wide application value in the fields of biological medicine, genetic engineering, disease treatment and DNA storage.
Owner:FUZHOU UNIV

Robust multi-sequence reconstruction method based on maximum a posteriori probability in DNA storage

The application discloses a robust multi-sequence reconstruction method based on maximum posterior probability in DNA storage, and comprises the following steps: decoding each sequence in a cluster by using an improved BCJR decoder; converting each sequence in the cluster into corresponding time sequence representation; determining the weight of each sequence; and deducing a MAP algorithm formula of robust multi-sequence decoding, and integrating the sequence weight into joint posterior probability calculation to inhibit the influence of outlying sequences. By proposing a base IDS channel model, the BCJR decoding algorithm is improved, so that the defects that the existing statistical inference sequence reconstruction method is only suitable for binary error correction code are overcome. Therefore, the application can process the quaternary error correction of the base form in the DNA sequence.
Owner:TIANJIN UNIV

DNA (deoxyribonucleic acid) data storage molecular tag based on nanopore as well as preparation method and application of DNA data storage molecular tag

The invention relates to a DNA data storage molecular tag based on nanopores and a preparation method and application thereof, and belongs to the technical field of DNA storage. The molecular tag is prepared by modifying azido polysaccharide onto an alkyne-containing DNA chain through a click chemical reaction. Hybridizing and fixing the molecular tag and a DNA bracket chain containing a specific notch structure, and scanning by using a glass nanopore; when the molecular tag passes through the nanopore, a characteristic ionic current blocking signal is generated, and accordingly data coding is achieved. The molecular tag disclosed by the invention is low in cost and simple and convenient to operate; polysaccharide with wide sources and efficient click chemistry are utilized, so that the preparation cost and complexity are remarkably reduced; the physicochemical properties of polysaccharide molecules are stable, the consistency of read signal amplitudes is ensured, the decoding accuracy is high, the label structure stability is good, and a guarantee is provided for long-term storage of data; based on a reversible hybridization storage structure, data updating can be realized without synthesizing a new DNA chain.
Owner:CHONGQING UNIV OF POSTS & TELECOMM +1

DNA storage encoding method based on graph convolution network and self-attention mechanism

The application discloses a DNA storage coding method based on a graph convolution network and a self-attention mechanism, and belongs to the technical field of coding in DNA storage. Specifically, a DNA coding sequence meeting a combination constraint condition is predicted, first, existing DNA coding is screened and data is cleaned to construct a DNA storage coding training set; second, a prediction model based on a graph convolution neural network and a self-attention mechanism is trained, and the self-attention mechanism is used to capture the relationship of local DNA coding; then, the coding data processed into a graph is input into the prediction model to perform coding prediction meeting the combination constraint; finally, a DNA storage coding set meeting the condition is output. The application constructs a DNA storage coding training set, trains a graph convolution self-attention neural network, better captures the relationship between codings, and adopts a learning-based prediction model to perform DNA storage coding, so that the application has high coding efficiency when processing codings with complex constraints.
Owner:DALIAN UNIV OF TECH

DNA sequencing read clustering method and system based on representation learning

The invention belongs to the crossing field of DNA digital storage and bioinformatics, and discloses a DNA sequencing read clustering method and system based on characterization learning, and the method comprises the steps: carrying out the preprocessing of an original DNA sequencing read, and obtaining a sequencing read and a corresponding variant 1 and variant 2 set; carrying out characterization learning on the DNA sequencing read segments through a deep learning model based on the sequencing read segments and the corresponding variant 1 and variant 2 sets; and based on the DNA sequencing read after characterization learning, realizing clustering of the DNA sequencing read through a fine tuning model. According to the method, the problem of clustering difficulty caused by sequencing errors in DNA storage is solved.
Owner:GUANGZHOU UNIVERSITY

A system and encoding and decoding method suitable for PB-level data DNA storage

The present application discloses a system and encoding / decoding method suitable for DNA storage of PB-level data. The data storage system includes multiple DNA clusters. The DNA cluster contains a data chain formed after binary data encoding, an XOR redundant chain generated by the data chain, an address chain of the DNA cluster at the storage system location, a basic information chain of the stored file, and a redundant chain generated by the XOR of the address chain and the file information chain. When a save operation is performed on a file with a preset suffix, a user-mode monitoring program receives an execution instruction and copies the file to a working directory, and encodes or decodes the file according to the file suffix. The result of the encoding or decoding operation is copied back to the original file directory. The present application enables the total capacity of the DNA database to reach above the PB level, allowing DNA storage to realize the application potential of high data density. At the same time, it can realize the automatic encoding of binary data and the automatic decoding of base sequences, streamline and automate the entire encoding and decoding process, and facilitate subsequent data management and archiving.
Owner:SHENZHEN INST OF ADVANCED TECH CHINESE ACAD OF SCI

DNA storage encryption and steganography method based on non-natural basic group

The invention discloses a DNA storage encryption and steganography method based on a non-natural basic group, and relates to the technical field of data storage. According to the method, a non-natural base pair (UBPs, NaM-TPT3) is introduced to a DNA sequence of coded information, during reading, the information sequence generates signal termination at the UBPs position, and a complete DNA sequence cannot be obtained, so that the purpose of encrypting DNA storage information is achieved, and DNA information can be smoothly read and decoded through bridge base isoTAT-mediated base conversion PCR (Polymerase Chain Reaction). Besides, a non-natural basic group is added to the 5'end of the index primer to prevent the index primer and a similar sequence from generating non-specific amplification, further, a natural primer is used for marking error information, a primer containing UBP is used for marking target information, and the target of information steganography is achieved through mixed storage. According to the method, encryption and steganography of DNA storage are realized by using a UBP technology, so that the security of DNA data storage is improved.
Owner:NORTHWESTERN POLYTECHNICAL UNIV

DNA storage method based on data superposition pseudorandom sequence index

The invention discloses a DNA (deoxyribonucleic acid) storage method based on data superposition pseudorandom sequence index, which comprises the following steps of: firstly, performing exclusive OR on a pseudorandom sequence and a sparse coding sequence to generate an addressable short fragment oligonucleotide sequence; then, by using the superposed pseudo-random sequence as an index, correlating the pseudo-random sequence hidden in the sequencing read with the known pseudo-random sequence at the boundary of each pseudo-random sequence by using a small sliding window to realize quick positioning of the read so as to determine the position of the read on the whole pseudo-random sequence; and finally, reconstructing a sparse coding sequence by adopting a majority voting algorithm, and realizing error-free recovery of data through sparse code decoding and iterative decoding. The method has the advantages that sequence positioning can be achieved only through correlation operation at the boundary of each pseudo-random sequence, local label sequences do not need to be added to oligonucleotide molecules, data recovery performance reduction caused by label sequence damage, sequence breakage and the like is avoided, and rapid data recovery can be achieved under low sequencing coverage.
Owner:TIANJIN UNIV SYNTHETIC BIOLOGY FRONTIER RES INST

Method and system for realizing DNA (Deoxyribonucleic Acid) storage by aiming at multi-rule rotation coding of Chinese text

The invention discloses a method and a system for realizing DNA storage by aiming at multi-rule rotation coding of Chinese texts, and relates to the technical field of DNA storage. Encoding the Chinese text by using a five-stroke font input method; mapping is carried out according to the positive and negative code tables, and an interval balance and dynamic detection strategy is introduced in the mapping process to control GC content distribution in intervals and the probability of occurrence of homopolymers between the intervals. Scattering and recombining the sequence by using block coding, and compressing by using RLE coding, wherein the RLE coding generates a sequence file and a run-length file; performing GC content constraint on the sequence file by using cross coding; and br compression is adopted for the run-length file to further improve the compression ratio. And generating a DNA sequence through rotary coding. According to the method, five-stroke coding, positive and negative code table mapping and sequence reconstruction strategies are introduced, so that the information storage density is higher, the local GC content of the DNA sequence is more stable, the homopolymer length is smaller, the unexpected motif proportion is lower, and the data has higher reliability and safety in the storage process.
Owner:DALIAN UNIV

Method and device for editing and reading edited information in DNA storage

The present application relates to the technical field of DNA storage, and more specifically to a method and apparatus for editing and reading information in DNA storage. The present application performs version identification and edit status identification on content information that requires editing, and encodes the version identification into version movable type and the edit status identification into edit status movable type, which are then encoded together with the content movable type encoding the content information into a storage movable type unit with a special storage structure. This innovative storage movable type unit not only allows for storage and reading of PNG images, GIF animations, TXT text, and MIDI music files, but also allows for dynamic editing of these content files. This provides a completely new method for editing DNA stored content information while ensuring data accuracy and storage stability.
Owner:WUHAN INST OF VIROLOGY CHINESE ACADEMY OF SCI

Iterative decoding method and system of VT code without prior information in DNA storage and storage medium

The application discloses a kind of DNA storage in prior information VT code iterative decoding method, system and storage medium, it is related to DNA storage technical field.The method is first obtained systematic code word sequence by VT code, generates noisy reception sequence by ID model simulation channel transmission;Again, joint grid chart is constructed and initialized, and soft information and bit decision information are obtained by completing SISO decoding calculation, and channel parameters are updated based on MAP drift tracking;After convergence by iterative control, multiple independent estimated code words are obtained, and finally original information sequence is recovered by multi-sequence majority voting fusion.The application constructs decoding-estimation closed-loop iteration framework, does not need prior channel parameters, can adaptively learn channel characteristics, capture error hot spot, improve decoding robustness by combining multi-sequence fusion, realize original data high reliability recovery, and promote DNA storage practicality.
Owner:TIANJIN UNIV

A normal-temperature stable sample DNA storage solution and a preparation method thereof

This invention discloses a preservation solution for preserving sample DNA at room temperature and its preparation method, belonging to the field of biological sample preservation technology. The preservation solution comprises: a lysis system, a nucleic acid protection system, and a buffering and stabilizing system; the lysis system comprises a complex surfactant and an enzymatic hydrolysis aid; the complex surfactant comprises: polyether polyol fatty acid ester and cocamidopropyl hydroxysulfonate betaine; the enzymatic hydrolysis aid comprises lysozyme Lyso-V and proteinase K; the nucleic acid protection system comprises: nitrogen-heterocyclic polyamine-carboxylic acid derivatives and dextran sulfate; the nitrogen-heterocyclic polyamine-carboxylic acid derivatives comprise: 1,4,7,10-tetraazacyclododecane-N,N',N'',N'''tetraacetic acid, 1,4,7-triazacyclononane-N,N',N''-triacetic acid, disodium ethylenediaminetetraacetate-nitrocyclic derivative, and diethylenetriaminepentaacetic acid-piperazine derivative; the buffering and stabilizing system comprises: an amphoteric buffer and a polymeric stabilizer; the amphoteric buffer comprises: 2-(N-morpholino)ethanesulfonic acid and N-tris(hydroxymethyl)methylglycine.
Owner:PEKING UNION MEDICAL COLLEGE HOSPITAL +1

End-to-end DNA data storage coding method and system based on multi-constraint loss

The invention discloses an end-to-end DNA data storage coding method and system based on multiple constraint loss, and relates to the technical field of DNA storage. A complementary matrix convolution detection mechanism is designed in a coding stage, a hairpin structure is found through the complementary matrix convolution detection mechanism, a micronizable loss function is constructed, and a micronizable hairpin structure constraint loss function is realized. An interpretable error predictor based on Transform is trained through real sequencing data, a penalty loss function is constructed by means of an attention mechanism error-prone motif mask matrix to achieve error-prone motif suppression, and then GC content loss and homopolymer loss are fused to form a multi-constraint loss function. According to the method, an end-to-end architecture with a Transform auto-encoder as a main body is constructed, mapping of image data and a DNA sequence is realized through differentiable arithmetic coding, data reconstruction is completed in combination with noise injection and reverse decoding, and in a training stage, a deep Q network is adopted to carry out adaptive joint adjustment on weights of a multi-constraint loss function and mean square error reconstruction loss, so that data reconstruction is completed. And end-to-end optimization of model parameters is realized.
Owner:DALIAN UNIV

DNA sequence direct encryption method and system based on automaton cryptography

The invention discloses a DNA sequence direct encryption method and system based on automaton cryptography, and relates to the technical field of DNA storage security. By designing a DS-mealy machine and a DS-cellular automaton, direct encryption of a DNA sequence is achieved, and the problems that in an existing DNA storage encryption method, compatibility with a coding model is poor, safety is insufficient, performance is low, and expandability is weak are solved. According to the method, sequence-level encryption is realized through base diffusion and rotation operation, so that base distribution is more uniform and random, and the security of DNA storage data is improved; meanwhile, the space resource consumption of encryption and decryption is reduced, the processing speed is increased, and the method is suitable for various DNA storage scenes and particularly suitable for storage of sensitive information such as medical data with high requirements for safety and efficiency.
Owner:DALIAN UNIV

A method and system for converting ancient text sequences for DNA storage data based on five-stroke coding

The application discloses a kind of ancient text sequence conversion method and system for DNA storage data based on five-stroke encoding, comprising the following steps: ancient text is converted into five-stroke encoding alphabet sequence according to word;Five-stroke encoding alphabet sequence is compressed by Huffman ternary, and ternary digital string is obtained;Ternary digital string is converted into DNA base sequence by dynamic mapping rule;DNA base sequence is segmented, and positioning code is added to each segment, and the final DNA sequence set suitable for DNA synthesis and high-throughput sequencing is output;The final DNA sequence set is synthesized, sequenced, and the ancient text is restored by reverse operation, and its reliability is verified.The method can efficiently and accurately convert ancient text to DNA storage sequence, ensure biological compatibility and data integrity restoration, and meet the long-term DNA storage needs of ancient text.
Owner:NANJING UNIV OF SCI & TECH