Counting Device Syntax Tree Subtree Matching
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing text mining devices cannot effectively count the occurrences of different expressions representing the same characteristic content across multiple input texts.
Innovation Solution
A counting device and method that analyzes and counts expressions by generating syntax trees, subtrees, and determining match relationships based on modifier-head relationships, allowing for the identification and counting of expressions across varying sentence structures.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Reliability
If text mining device uses expression conversion to find texts with same characteristic content, then text association capability is improved, but expression usage counting capability deteriorates
Solution Approach 1:
The patent segments expressions into subtrees with modifier-head relationships, allowing individual counting of each expression variant while maintaining their semantic associations. This segmentation enables both text mining functionality and expression counting capability to coexist.
Solution Approach 2:
The patent introduces an intermediary counting mechanism that operates between the expression conversion process and the final text association results. This intermediary layer tracks expression usage frequencies without interfering with the text mining functionality.
2Ease of operation
If text mining device converts expressions to prescribed expressions, then text normalization is improved, but original expression information is lost
Solution Approach 1:
The patent creates copies of the original expressions in subtree form during the conversion process, preserving the original expression information while enabling normalized text processing. Each subtree represents a copy of the original expression structure that can be counted and analyzed.
3Adaptability or versatility
If counting device analyzes multiple input texts with different expressions, then expression variety coverage is improved, but counting accuracy deteriorates due to structural variations
Solution Approach 1:
The patent applies local quality by analyzing each subtree's modifier-head relationships individually, allowing accurate counting of each expression variant while accommodating structural variations across different texts. Each local subtree structure is evaluated on its own merits.
Data Source
Figure 1
Figure 2
Figure 3
AI summary
A counting device (100) provided with a subtree generating part (123) for generating first subtree comprising a first sentence and a second subtree comprising a second sentence. The counting device (100) is provided with: a categorizing part (125) for categorizing the first subtree in the same group as the second subtree when it is determined that a first expression represented by the first subtree and a second expression represented by a second subtree represent a matching content; and an output part (127) for outputting the number of subtrees categorized in the group, or an expression represented by a plurality of syntax trees or one of the subtrees categorized in the aforementioned group.