Display Device Voice Search Intent Analysis
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing voice recognition search systems fail to process user queries with ambiguous speech that does not include exact content metadata words, leading to incorrect search results.
Innovation Solution
A display device equipped with a voice acquisition unit and a processor that performs intention analysis on user speech, allowing for successful content metadata search and display even with ambiguous input, using a combination of Speech To Text and natural language processing engines.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Measurement precision
If exact program, character, or genre names are used for voice search, then search accuracy is improved, but ease of operation deteriorates
Solution Approach 1:
The patent introduces an intermediary processing layer between user speech and content metadata matching. This layer includes speech-to-text conversion, natural language processing, and semantic analysis components that translate ambiguous user speech into structured search queries, enabling accurate content retrieval without requiring exact metadata terminology
Solution Approach 2:
The system changes the parameter of speech processing from exact keyword matching to semantic meaning analysis. By transforming the search mechanism from requiring precise metadata term matching to understanding user intent through natural language processing, the system allows flexible, ambiguous speech while maintaining high search accuracy
2Measurement precision
If voice recognition search uses exact metadata matching, then search precision is improved, but adaptability deteriorates
Solution Approach 1:
The patent implements a universal speech processing framework that handles multiple types of user queries (exact titles, descriptions, character names, plot summaries) through a single integrated system. The natural language processing engine can interpret various speech patterns and convert them into appropriate search operations, making the system adaptable to different user needs while maintaining precision
Solution Approach 2:
The intermediary natural language processing layer acts as a bridge between diverse user speech patterns and the content metadata database. It translates various forms of ambiguous speech into structured search queries that can accurately retrieve relevant content, thereby enhancing both adaptability and search precision simultaneously
Data Source
AI summary
The present disclosure provides a display device comprising: a voice acquisition unit having at least one microphone for acquiring user speech; and a processor for acquiring text data corresponding to the user speech, acquiring intention analysis results by performing intention analysis on the basis of the text data, determining whether the intention analysis is successful on the basis of the intention analysis results, searching for content metadata corresponding to the intention analysis results on the basis of the intention analysis results if the intention analysis is successful, and displaying, through a display unit, content search results corresponding to the user speech on the basis of the retrieved content metadata.


