system

JP2026068480APending Publication Date: 2026-04-22SOFTBANK GROUP CORP
View PDF 1 Cites 0 Cited by

Patent Information

Authority / Receiving Office
JP · JP
Patent Type
Applications
Current Assignee / Owner
SOFTBANK GROUP CORP
Filing Date
2024-10-10
Publication Date
2026-04-22

AI Technical Summary

Benefits of technology

【0005】 [本発明は、利用者の音声を受信し、音声データとして取得する手段を設けることにより開始する。次に、この音声データを解析し、テキストデータに変換する音声認識技術を提供する手段を含む。さらに、利用者の口の動きを捕捉し、動作データを生成する動作認識技術を有する手段を備え、音声認識の精度を向上させる。また、これらのデータをサーバに送信し、サーバ内でテキストデータを異なる言語に翻訳し、翻訳されたテキストを音声データに変換する手段を提供する。最後に、生成された音声データを利用者の端末に送信し、再生する手段を含むことで、自然で個別性の高い音声翻訳を実現する]。

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure 2026068480000001_ABST
    Figure 2026068480000001_ABST
Patent Text Reader

Abstract

We provide a system that enables more natural and accurate speech translation. [Solution] The system includes means for receiving the user's voice and acquiring audio data, means for providing speech recognition technology that analyzes the audio data and converts it into text data, means for motion recognition technology that captures the user's mouth movements and generates motion data, means for transmitting the text data and motion data to a server, means for the server to translate the text data into different languages ​​and convert the text in the different languages ​​into audio data, and means for transmitting the audio data to the user's terminal and playing it back.
Need to check novelty before this filing date? Find Prior Art

Claims

1. A means of receiving the user's voice and acquiring audio data, A means for providing speech recognition technology that analyzes the aforementioned audio data and converts it into text data, Means having motion recognition technology that captures the mouth movements of the user and generates motion data, Means for transmitting the aforementioned text data and operational data to a server, A server provides means for translating the text data into different languages ​​and converting the text in the different languages ​​into audio data. A system including means for transmitting the aforementioned audio data to a user's terminal and playing it back.

2. The system according to claim 1, wherein the server maintains the user's voice characteristics and linguistic features and reflects them in the generation of translated voice data.

3. The system according to claim 1, wherein the motion recognition technology is used to improve the accuracy of speech recognition based on the user's mouth movements.

Citation Information

Patent Citations

  • Persona chatbot control method and system

    JP2022180282A