Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

5 results about "Sequence annotation" patented technology

Sequence annotation. Sequence annotation is the "process of marking specific features in a DNA, RNA or protein sequence with descriptive information about structure or function".

Bidirectional reversible conversion method and system between peptide molecule SMILES and sequence expression

The invention discloses a bidirectional reversible conversion method and system between a peptide molecule SMILES and a sequence expression. The core innovation lies in that a new sequence description syntax is defined to retain information of a polypeptide special bond and specific modification of amino acid; a main chain atom index and adjacency traversal topology identification algorithm is adopted, and end group and topology integrated detection and coding are carried out; a residue recognition algorithm for main chain cutting and template library matching is compatible with any standard or non-standard amino acid residues, an extensible end group library / monomer template library and an automatic increment mechanism, and automatic recognition and sequence annotation of S-S disulfide bonds; the invention relates to a high-fidelity assembly algorithm of HELM anchor points and topology aware cyclic peptide processing. The method solves the problems of incapability of supporting a complex polypeptide topological structure, poor reversibility, insufficient expansibility of a monomer library and the like in the prior art, can be widely applied to scenes of quantitative structure-activity relationship model construction, large-scale polypeptide data cleaning and the like, and has remarkable practicability and innovativeness.
Owner:ANGXIN BIOTECHNOLOGY CO LTD

Bidding document structured information extraction and intelligent bid evaluation integrated system

PendingCN122334227AEngineeringSequence annotation
The present application relates to the technical field of bid evaluation system, in particular to a bid document structured information extraction and intelligent bid evaluation integrated system, which comprises a rule analysis module, a constraint calculation module, a field construction module, an element integration module and a consistency verification module.In the present application, semantic boundary recognition and sequence annotation are performed on the scoring clause text to make the scoring condition expression structured and analyzable, a constraint parameter strength analysis is introduced to form a weight sequence, different conditions are presented with a distinguishing degree in the review, the bid document field comparison is focused on high correlation content, and the element correlation relationship is constructed in combination with the paragraph and chapter position information, the consistency and conflict review is carried out at the bid evaluation element level, the redundant field influence is compressed, the scoring basis integrity and logical coherence are enhanced, the review process expression is clearer, and the result formation path is more checkable and stable.
Owner:GUANGDONG POWER GRID CO LTD INFORMATION CENT

Method, device and medium for constructing barcode gene database for eDNA analysis

The invention discloses a method and equipment for constructing a bar code gene database for eDNA analysis and a medium, and belongs to the technical field of gene database construction. According to the method, a standardized preprocessing process is deployed for each data source, and a taxonomy mapping system based on same-object different-name analysis and multi-level backtracking inference is combined, so that the sequence annotation success rate and integration efficiency of the heterogeneous data sources are improved, high fault-tolerant analysis of non-standard and incomplete taxonomy names is realized, and data waste is reduced. Meanwhile, taxonomy information synchronous updating and traceability checking are constructed, batch optimization and correction of historical data are supported while the quality of newly added data is guaranteed, and long-term accuracy and scientificity of a database are guaranteed.
Owner:ONE HEALTH BIOTECHNOLOGY (SUZHOU) CO LTD

Two-way reversible conversion method and system between peptide molecular smiles and sequence list expression

The application discloses a bidirectional reversible conversion method and system between peptide molecules SMILES and sequence expressions, and the core innovation is that: a new sequence description grammar is defined to retain the information of special bonds of polypeptides and specific modifications of amino acids; a topological identification algorithm of main chain atom index and adjacency traversal, end group and topological integration detection and coding; a residue identification algorithm compatible with any standard or non-standard amino acid residue of main chain cutting and template library matching, an expandable end group library / monomer template library and an automatic increment mechanism, automatic identification and sequence annotation of S-S disulfide bond; a high-fidelity assembly algorithm of HELM anchor points and topologically-aware cyclic peptide processing. The application solves the problems that the prior art cannot support complex polypeptide topological structure, has poor reversibility, and has insufficient monomer library expansion, and can be widely applied to quantitative structure-activity relationship model construction, large-scale polypeptide data cleaning and other scenes, and has obvious practicability and innovation.
Owner:ANGXIN BIOTECHNOLOGY CO LTD

Event argument detection method and system based on label sequence consistency modeling

This invention proposes an event argument detection method and system based on label sequence consistency modeling. It mainly includes word sequence semantic encoding, word label sequence annotation, error-prone label sequence generation, and contrastive learning regularization. Word sequence semantic encoding uses BERT and a trained language model to learn semantic representations of preprocessed words, incorporating event type information into the representation vector. Word label sequence annotation uses a fully connected network to predict the probability distribution of the label corresponding to each word. Error-prone label sequence generation generates error-prone label sequences according to a certain strategy based on the probability distribution of the word label sequences. Contrastive learning regularization constructs a regularization loss based on the contrastive learning between error-prone label sequences and correct label sequences, improving the consistency of word sequence labels.
Owner:INST OF COMPUTING TECH CHINESE ACAD OF SCI