Reduced Computation Hidden Markov Model for Genomic Variant Calling

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Current bioinformatics methods for constructing genomic sequences and determining variants are labor-intensive, time-consuming, and prone to errors due to the complexity of processing large DNA sequence fragments and the presence of noise and error sources in sequencing data.

Innovation Solution

A reduced computation hidden Markov model (HMM) method and system that pre-filters and processes genomic data using a HMM pre-filter engine to generate control parameters, selecting a reduced number of HMM matrix cells for computation, thereby accelerating the variant calling process and improving accuracy.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Measurement precision

If a full HMM matrix computation is performed on all cells, then the variant calling accuracy is maintained, but the computational time and resources increase significantly

Engineering Contradiction:
Improvevariant calling accuracyVSAvoidcomputational time
Core Design Contradiction:
Measurement precisionVSLoss of time

Solution Approach 1:

The HMM matrix is segmented into multiple blocks or regions, where only the most relevant blocks are computed in full detail while less critical blocks use approximations or are skipped entirely. This allows the system to maintain accuracy for the most important variant calls while reducing overall computational burden.

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

The system performs partial computation of the HMM matrix by calculating only the necessary portions required for accurate variant calling. By identifying and computing only the critical cells that contribute most to the final accuracy, the system achieves sufficient precision without the excessive computational cost of full matrix evaluation.

Inventive Principle:
Principle #16Partial or excessive action

2Reliability

If all HMM matrix cells are computed, then comprehensive variant detection is achieved, but the computational complexity increases

Engineering Contradiction:
Improvevariant detection comprehensivenessVSAvoidcomputational complexity
Core Design Contradiction:
ReliabilityVSDevice complexity

Solution Approach 1:

The system performs preliminary filtering and preprocessing steps before the main HMM computation, identifying and prioritizing the most likely variant regions. This preliminary action allows the subsequent HMM computation to focus only on high-probability areas, reducing overall computational complexity while maintaining detection comprehensiveness.

Inventive Principle:
Principle #10Preliminary action

Solution Approach 2:

Different regions of the HMM matrix are treated with different levels of computational quality. High-priority regions receive full computational treatment to ensure reliable variant detection, while lower-priority regions use simplified models or are excluded from computation, thereby reducing overall complexity while maintaining comprehensive detection capability.

Inventive Principle:
Principle #3Local quality

3Measurement precision

If the complete HMM computation process is used, then accurate genomic sequence construction is achieved, but the processing speed decreases

Engineering Contradiction:
Improvegenomic sequence accuracyVSAvoidprocessing speed
Core Design Contradiction:
Measurement precisionVSProductivity

Solution Approach 1:

The system dynamically adjusts the level of HMM computation based on real-time data characteristics, read quality scores, and regional importance. This dynamic approach allows the system to maintain high accuracy for critical regions while using faster, simplified methods for less critical areas, thereby improving overall processing speed without sacrificing genomic sequence accuracy.

Inventive Principle:
Principle #15Dynamics

Solution Approach 2:

The system changes computational parameters such as matrix resolution, computation depth, and precision levels based on the specific requirements of different genomic regions. By adjusting these parameters dynamically, the system achieves accurate genomic sequence construction where needed while maintaining high processing speed in less critical areas.

Inventive Principle:
Principle #35Parameter changes

Data Source

PatentUS20220230084A1Method and System for a Reduced Computation Hidden Markov Model in Computational Biology Applications
Publication Date: 2022.07.21 ILLUMINA INC
  • US20220230084A1 patent drawing
  • US20220230084A1 patent drawing
  • US20220230084A1 patent drawing

AI summary

A system and method for a reduced computation hidden markov model (HMM) in computational biology applications is disclosed herein. The method includes performing a correlation between the haplotype sequence and the read sequence at the HMM pre-filter engine. The method includes computing a MINI metric from a reduced number of cells at the HMM computation engine.