一种关系型数据库基数估计处理方法、装置以及设备

By dividing attribute subsets based on attribute relationships in a relational database and performing frequency distribution statistics and multivariate function modeling, the problem of inaccurate cardinality estimation caused by insufficient data relationships is solved, thereby improving the accuracy of cardinality estimation and the reliability of query plans.

CN116361326BActive Publication Date: 2026-07-17NORTHEASTERN UNIV CHINA +1

Patent Information

Authority / Receiving Office
CN · China
Patent Type
Patents(China)
Current Assignee / Owner
NORTHEASTERN UNIV CHINA
Filing Date
2023-03-13
Publication Date
2026-07-17

AI Technical Summary

Technical Problem

Existing cardinality estimation schemes do not adequately consider data correlation, which affects the accuracy of cardinality estimation results and consequently the reliability of query plans or optimization schemes in relational databases.

Method used

By determining the set of attributes in the database table, dividing it into multiple attribute subsets based on the relationships between attributes, performing horizontal and vertical segmentation, statistically analyzing the frequency distribution, and using multivariate function modeling, cardinality estimation is performed.

Benefits of technology

It improves the accuracy of cardinality estimation, enhances the reliability of selecting query plans or optimization schemes, and improves query efficiency.

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure CN116361326B_ABST
    Figure CN116361326B_ABST
Patent Text Reader

Abstract

本说明书实施例公开了关系型数据库基数估计处理方法、装置以及设备。包括:确定指定的数据库表所包含的第一属性集合;根据第一属性集合内的属性之间的关联性,从第一属性集合内划分出多个属性子集合,包括子集合内属性关联性弱的第二属性子集合、子集合内属性关联性强且与第二属性子集合关联性强的第三属性子集合、子集合内属性关联性强且与其他属性子集合关联性弱的第四属性子集合;通过对第二属性子集合进行横向和 / 或纵向的切分,进行频率分布统计;从第三属性子集合内,划分出多个与第二属性子集合关联性相对弱的属性次级子集合,将属性次级子集合和第四属性子集合建模为多变量函数;根据频率分布统计结果和多变量函数,进行基数估计。
Need to check novelty before this filing date? Find Prior Art