Technology for responding to remarks using speech synthesis
a technology of speech synthesis and response voice, which is applied in the field of speech or voice synthesis apparatus and system, can solve the problem that the voice output of voice synthesis gives the user an unnatural feeling, and achieve the effect of easy and controllable replying voi
Patent Information
- Authority / Receiving Office
- US · United States
- Patent Type
- Applications(United States)
- Current Assignee / Owner
- Publication Date
- 2017-04-20
Smart Images

Figure 1 
Figure 2 
Figure 3
Abstract
Description
TECHNICAL FIELD
[0001] The present invention relates to a speech or voice synthesis apparatus and system which, in response to a remark, question or utterance made by voice input, provide replying output, as well as a coding / decoding device related to the voice synthesis.BACKGROUND ART
[0002] In recent years, the following voice synthesis techniques have been proposed. Examples of such proposed voice synthesis techniques include a technique that synthesizes and outputs voice corresponding to a speaking tone and voice quality of a user and thereby generates voice in a more human-like manner (see, for example, Patent Literature 1), and a technique that analyzes voice of a user to diagnose psychological and health states etc. of the user (see, for example, Patent Literature 2).
[0003] Also proposed in recent years is a voice interaction or dialogue system which implements voice interaction with a user by outputting, in synthesized voice, content designated by a scenario while recognizing voi...
Examples
first embodiment
[0102]First of all, a first embodiment of a voice synthesis apparatus of the present invention will be described. FIG. 1 is a block diagram showing a construction of the first embodiment of the voice synthesis apparatus 10 of the present invention. In FIG. 1, the voice synthesis apparatus 10 is a terminal apparatus, such as a mobile or portable apparatus, including a CPU (Central Processing Unit), a voice input section 102, and a speaker 142. In the voice synthesis apparatus 10, a plurality of functional blocks are built as follows by the CPU executing a preinstalled application program.
[0103]More specifically, in the tone synthesis apparatus 10 are built a voice-utterance-section detection section 104, a pitch analysis section 106, a linguistic analysis section 108, a reply creation section 110, a voice synthesis section 112, a linguistic database 122, a reply database 124, an information acquisition section 126 and a voice library 128. Namely, each of the functional blocks in the ...
second embodiment
[0133]The following describe a second embodiment of the voice synthesis apparatus 10 of the present invention, which employs a modification of the replying voice generation method. FIG. 8 is a block diagram showing a construction of the second embodiment of the voice synthesis apparatus 10 of the present invention. Whereas the above-described first embodiment is constructed in such a manner that the reply creation section 110 outputs a voice sequence where a pitch is allocated per sound (syllable) of a replying language responsive to a question and that the voice synthesis section 112 synthesizes voice of a reply (replying voice) on the basis of the voice sequence, the second embodiment is constructed in such a manner that the replying voice output section 113 acquires a reply (response) to a question (remark) and generates and outputs voice waveform data of the entire reply (response).
[0134]Examples of the above-mentioned reply (response) include one created by the replying voice o...
application examples
and Modifications
[0138]It should be appreciated that the present invention is not limited to the above-described first and second embodiments and various other application examples and modifications of the present invention are also possible as follows. Further, any selected ones of the plurality of application examples and modifications may be combined as appropriate.
[0139]
[0140]Whereas the embodiments of the invention have been described above in relation to the case where the voice input section 102 inputs user's voice (remark) via the microphone and converts the input voice (remark) into a voice signal, the present invention is not so limited, and the voice input section 102 may be configured to receive a voice signal, processed by another processing section or supplied (or forwarded) from another device, via a recording medium, a communication network or the like. Namely, the voice input section 102 may be configured in any desired manner as long as it receives an input voice s...