Intelligent Screen Reading for Context- and Emotion-Aware Audio

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Existing electronic devices lack the ability to intelligently read displayed content, failing to associate intent, context, and emotion, leading to confusion for visually impaired and general users.

Innovation Solution

An electronic device equipped with an intelligent screen reading engine that analyzes displayed content to extract insights on intent, importance, emotion, and sound representation, using deep neural networks and generative text reading to provide meaningful audio output.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Loss of information

If existing text-to-speech method is used to read displayed content, then the device can read aloud text and emoji definitions, but the reading lacks emotional meaning and context understanding

Engineering Contradiction:
Improveemotional meaning and contextVSAvoidreading engine complexity
Core Design Contradiction:
Loss of informationVSDevice complexity

Solution Approach 1:

The patent introduces an intelligent screen reading engine as an intermediary between the displayed content and the user. This engine analyzes the visual content, extracts emotional meaning and context, and generates enhanced audio output with appropriate emotional tone, thereby recovering the lost emotional information without requiring complete system redesign

Inventive Principle:
Principle #24Intermediary (Mediator)

Solution Approach 2:

The patent replaces the mechanical text-to-speech reading system with an intelligent system that uses deep learning models (BERT, GPT) to understand and generate emotionally nuanced audio. This substitution transforms the reading process from simple text conversion to intelligent content comprehension and emotional expression

Inventive Principle:
Principle #28Mechanics substitution (Replace mechanical system)

2Loss of information

If the device reads every displayed content element in detail, then complete information is provided, but the user gets confused and loses the actual intent

Engineering Contradiction:
Improvecontent completenessVSAvoiduser understanding
Core Design Contradiction:
Loss of informationVSEase of operation

Solution Approach 1:

The intelligent screen reading engine extracts only the most relevant and meaningful elements from the displayed content, separating essential information from redundant details. It identifies key semantic units, emotional indicators, and contextual importance, presenting only what matters to the user while maintaining complete understanding of the original content

Inventive Principle:
Principle #2Taking out (Extraction)

Solution Approach 2:

The patent applies different processing qualities to different parts of the displayed content based on their importance. Critical information receives detailed analysis and prominent audio presentation, while less important elements receive simplified handling. This local differentiation optimizes both information completeness and user comprehension

Inventive Principle:
Principle #3Local quality

3Productivity

If the device reads content without understanding meaning and intent, then the reading process is simple and fast, but the output appears mechanical and lacks human-like emotion

Engineering Contradiction:
Improvereading speedVSAvoidintent and context understanding
Core Design Contradiction:
ProductivityVSLoss of information

Solution Approach 1:

The intelligent screen reading engine performs preliminary analysis of the displayed content before generating audio output. It pre-processes the visual information to extract semantic meaning, emotional tone, and contextual relationships, preparing the groundwork for emotionally accurate reading while maintaining efficient processing throughput

Inventive Principle:
Principle #10Preliminary action

Data Source

PatentUS12525219B2Method and electronic device for intelligently reading displayed contents
Publication Date: 2026.01.13 SAMSUNG ELECTRONICS CO LTD
  • US12525219B2 patent drawing
  • US12525219B2 patent drawing
  • US12525219B2 patent drawing

AI summary

A method for intelligently reading displayed contents by an electronic device is provided. The method includes obtaining a screen representation based on a plurality of contents displayed on a screen of the electronic device. The method includes extracting a plurality of insights comprising at least one of intent, importance, emotion, sound representation and information sequence of the plurality of contents from the plurality of contents based on the screen representation. The method includes generating audio emulating the extracted plurality of insights.