Object Information Translation Device for Non-Text Content

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Existing machine translation technologies only translate text inputs and fail to fully meet user requirements as they do not account for the translation of objects, leading to incomplete understanding of unfamiliar items like foreign drugs or commodities without accompanying text, and cannot handle non-text content such as billboards encountered during travel.

Innovation Solution

A method and device for translating object information, which recognizes source objects using multimedia data and provides target-object information, including derivative information, to enhance understanding and meet user needs, even in extreme conditions like constrained networks or incomplete input.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Adaptability or versatility

If existing machine translation technologies translate only text inputs, then translation efficiency is maintained, but the applicability and completeness of translation fail to meet user requirements

Engineering Contradiction:
Improveapplicability of translationVSAvoidcomplexity of translation system
Core Design Contradiction:
Adaptability or versatilityVSDevice complexity

Solution Approach 1:

The translation system is extended to handle multiple types of input beyond text, including images, audio, and video content. The system performs translation on various media types by extracting text from these inputs and applying translation models, making the system universally applicable to diverse translation needs while maintaining a unified processing framework

Inventive Principle:
Principle #6Universality (Multi-functionality)

2Loss of information

If machine translation translates only text, then the translation process remains simple, but user understanding of unfamiliar items like foreign drugs or commodities is incomplete

Engineering Contradiction:
Improvecompleteness of translation informationVSAvoidcomplexity of information processing
Core Design Contradiction:
Loss of informationVSDevice complexity

Solution Approach 1:

The system performs preliminary extraction of text from various media types (images, audio, video) before applying translation. By pre-processing the input to extract translatable text content, the system ensures that no translation information is lost while keeping the overall process manageable through staged processing

Inventive Principle:
Principle #10Preliminary action

3Adaptability or versatility

If the system translates only text inputs, then processing speed is maintained, but the system cannot handle non-text content such as billboards encountered during travel

Engineering Contradiction:
Improverange of translation subjectsVSAvoidtranslation processing efficiency
Core Design Contradiction:
Adaptability or versatilityVSProductivity

Solution Approach 1:

The translation system segments different types of input content (text, images, audio, video) and processes each type through appropriate extraction and translation pipelines. This segmentation allows the system to handle diverse translation subjects like billboards and signs while maintaining efficient processing by routing each input type through its optimized path

Inventive Principle:
Principle #1Segmentation

Data Source

PatentUS10990768B2Method and device for translating object information and acquiring derivative information
Publication Date: 2021.04.27 SAMSUNG ELECTRONICS CO LTD
  • US10990768B2 patent drawing
  • US10990768B2 patent drawing
  • US10990768B2 patent drawing

AI summary

A method and device are provided for translating object information and acquiring derivative information, including obtaining, based on the acquired source-object information, target-object information corresponding to the source object by translation, and outputting the target-object information. A language environment corresponding to the source object is different from a language environment corresponding to the target object. By applying the present disclosure, the range of machine translation subjects can be expanded, and the applicability of translation can be enhanced, a user's requirements on translation of objects can be met.