Voice playback method, device, equipment and storage medium

By displaying the tone customization prompt page in the terminal system, obtaining the audio data uploaded by the user and generating the target tone configuration information, the problem of non-interoperability of tone settings between applications in the prior art is solved, and the effect of setting tone for terminal devices is achieved quickly and conveniently.

CN114121028BActive Publication Date: 2025-05-13TENCENT TECHNOLOGY (SHENZHEN) CO LTD
View PDF 1 Cites 0 Cited by

Patent Information

Application Number
CN202111137785.3
Authority / Receiving Office
CN · China
Patent Type
Patents(China)
Current Assignee / Owner
Filing Date
2021-09-27
Publication Date
2025-05-13
Estimated Expiration
2041-09-27

AI Technical Summary

Technical Problem

The existing voice synthesis technology can only set the tone for a single application, resulting in the tone settings between applications being incompatible, and the user setting the tone process is cumbersome.

Method used

By displaying the tone customization prompt page in the terminal system, obtaining the audio data uploaded by the user, generating the target tone configuration information, and responding to user settings, setting the tone for the terminal device quickly and conveniently.

Benefits of technology

It realizes the quick and convenient setting of tones for terminal devices, improves user experience and is highly applicable.

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure CN114121028B_ABST
    Figure CN114121028B_ABST
Patent Text Reader

Abstract

The embodiments of the present application disclose a voice playback method, device, equipment and storage medium, which can be applied to various scenarios such as cloud technology, artificial intelligence, smart transportation, Internet of Things, assisted driving, etc. The method includes: in response to the terminal system of the user logging into the target terminal, displaying the timbre customization prompt page; obtaining the first audio data uploaded by the user based on the timbre customization prompt page, displaying the timbre list page, the timbre list page includes the first timbre configuration information determined by the first audio data, and the first audio data and the first timbre configuration information correspond to the same timbre; in response to the user's setting instruction for the target timbre configuration information in the timbre list page, the audio information is played through the target terminal with the timbre corresponding to the target timbre configuration information. By adopting the embodiments of the present application, the timbre can be set for the terminal quickly and conveniently, and it is highly applicable.
Need to check novelty before this filing date? Find Prior Art

Description

Technical Field

[0001] The present application relates to the field of Internet of Things, and in particular to a voice playback method, device, equipment and storage medium. Background Art

[0002] With the development of science and technology, speech synthesis technology has made great progress, and machine voice broadcasting has been widely used in smart mobile terminals, smart homes, car audio and other devices.

[0003] However, existing speech synthesis technologies are often only targeted at a single application, that is, developers of different applications only provide independent timbre synthesis solutions for their respective applications. Taking smart mobile terminals as an example, users can only set specific timbres for individual applications, which results in the timbre settings between applications not being interoperable, and users cannot set specific timbres for smart mobile terminals. And if users need to set a specific timbre for a smart mobile terminal, they need to set the same timbre for each application, which is a cumbersome process. Therefore, how to quickly and conveniently set the timbre for a terminal has become an urgent problem to be solved. Summary of the invention

[0004] The embodiments of the present application provide a voice playback method, apparatus, device, and storage medium, which can quickly and conveniently set the timbre for a terminal device and have high applicability.

[0005] On the one hand, an embodiment of the present application provides a voice playback method, the method comprising:

[0006] In response to the user logging into the terminal system of the target terminal, displaying a tone customization prompt page;

[0007] Acquire the first audio data uploaded by the user based on the timbre customization prompt page, and display a timbre list page, wherein the timbre list page includes the target timbre configuration information determined by the first audio data, and the first audio data and the target timbre configuration information correspond to the same timbre;

[0008] In response to the user's setting instruction for the target timbre configuration information in the timbre list page, the target terminal plays the audio information with the timbre corresponding to the target timbre configuration information.

[0009] On the other hand, an embodiment of the present application provides a voice playback device, the device comprising:

[0010] A prompt page display module, used for displaying a tone customization prompt page in response to a user logging into a terminal system of a target terminal;

[0011] A timbre list display module, used to obtain the first audio data uploaded by the user based on the timbre customization prompt page, and display a timbre list page, wherein the timbre list page includes the target timbre configuration information determined by the first audio data, and the first audio data and the target timbre configuration information correspond to the same timbre;

[0012] The voice playing module is used to respond to the setting instruction of the user for the target timbre configuration information in the timbre list page, and play the audio information with the timbre corresponding to the target timbre configuration information through the target terminal.

[0013] On the other hand, an embodiment of the present application provides an electronic device, including a processor and a memory, wherein the processor and the memory are connected to each other;

[0014] The memory is used to store computer programs;

[0015] The above-mentioned processor is configured to execute the voice playback method provided in the embodiment of the present application when calling the above-mentioned computer program.

[0016] On the other hand, an embodiment of the present application provides a computer-readable storage medium, which stores a computer program, and the computer program is executed by a processor to implement the voice playback method provided in the embodiment of the present application.

[0017] On the other hand, an embodiment of the present application provides a computer program product, which includes a computer program or computer instructions. When the above computer program or computer instructions are executed by a processor, the voice playback method provided in the embodiment of the present application is provided.

[0018] In the embodiment of the present application, after the user logs in to the terminal system of the target terminal, the audio data uploaded by the user can be obtained based on the tone customization prompt page, so that the tone can be customized based on the audio data uploaded by the user to improve the user experience. And based on the user's setting instructions for each tone configuration information in the tone list page, the tone can be customized for the target terminal quickly and conveniently, with high applicability. BRIEF DESCRIPTION OF THE DRAWINGS

[0019] In order to more clearly illustrate the technical solutions in the embodiments of the present application, the drawings required for use in the embodiments will be briefly introduced below. Obviously, the drawings described below are only some embodiments of the present application. For ordinary technicians in this field, other drawings can be obtained based on these drawings without paying creative work.

[0020] Figure 1 It is a flowchart of the voice playback method provided in the embodiment of the present application;

[0021] Figure 2This is a schematic diagram of a login page of a terminal system provided in an embodiment of the present application;

[0022] Figure 3 It is a scene schematic diagram of the login page of the tone recording program provided in an embodiment of the present application;

[0023] Figure 4 is a scene diagram of a tone recording page provided in an embodiment of the present application;

[0024] Figure 5a is a scene schematic diagram of the first guide page provided in an embodiment of the present application;

[0025] Figure 5b is a schematic diagram of a scenario of a second guide page provided in an embodiment of the present application;

[0026] Figure 5c is a schematic diagram of a scenario of a third guide page provided in an embodiment of the present application;

[0027] Figure 6 is a schematic diagram of a scenario of an audio information improvement page provided in an embodiment of the present application;

[0028] Figure 7a This is a schematic diagram of a scene of a tone list page provided in an embodiment of the present application;

[0029] Figure 7b This is another schematic diagram of a timbre list page provided in an embodiment of the present application;

[0030] Figure 8 is a scene schematic diagram of a tone setting page provided in an embodiment of the present application;

[0031] Fig. 9 It is a schematic diagram of the process of customizing the sound quality of the vehicle terminal provided in the embodiment of the present application;

[0032] Fig.10a It is a functional framework diagram of the TTS component provided in the embodiment of the present application;

[0033] Fig.10b It is a schematic diagram of the process of using the TTS component provided in the embodiment of the present application;

[0034] Fig.11a This is a timing diagram of TTS service selection provided by an embodiment of the present application;

[0035] Fig.11b is a timing diagram of setting the tone provided in an embodiment of the present application;

[0036] Fig.12 is a structural diagram of a voice playback device provided in an embodiment of the present application;

[0037] Fig.13 It is a schematic diagram of the structure of an electronic device provided in an embodiment of the present application. DETAILED DESCRIPTION

[0038] The following will be combined with the drawings in the embodiments of the present application to clearly and completely describe the technical solutions in the embodiments of the present application. Obviously, the described embodiments are only part of the embodiments of the present application, not all of the embodiments. Based on the embodiments in the present application, all other embodiments obtained by ordinary technicians in this field without creative work are within the scope of protection of this application.

[0039] The voice playback method provided in the embodiment of the present application can be applied to relevant scenarios involving voice broadcast technology such as the Internet of Things, the Internet of Vehicles, and artificial intelligence. The specific application can be determined based on the actual application scenario requirements and is not limited here. For example, based on the voice playback method provided in the embodiment of the present application, the control of the broadcast tone of the vehicle terminal in the Internet of Vehicles and the device control terminal in the Internet of Things (such as smart home devices) can be realized, which has high applicability.

[0040] Among them, the voice playback method provided in the embodiment of the present application can be executed by a server, a TTS (Text To Speech) component or a terminal, which can be determined based on the actual application scenario requirements and is not limited here. The above-mentioned server can be an independent physical server, such as an Internet of Vehicles server, or a server cluster or distributed system composed of multiple physical servers, or a cloud server that provides cloud computing services, which is not limited in this application. The above-mentioned terminal can be a smart phone, a tablet computer, a laptop computer, a desktop computer, a smart speaker, a smart watch, a car terminal, a smart TV, etc., but is not limited to this.

[0041] See also Figure 1 , Figure 1 is a flow chart of the voice playback method provided in the embodiment of the present application. Figure 1 As shown, the voice playback method provided in the embodiment of the present application may include the following steps:

[0042] Step S11: In response to the user logging into the terminal system of the target terminal, a tone customization prompt page is displayed.

[0043] In some feasible implementations, the target terminal is a terminal that broadcasts voice to the user, such as a vehicle terminal (i.e., a vehicle-mounted terminal), etc. After the target terminal is started, the target terminal can display a login page for logging into the terminal system of the target terminal to the user, such as a QR code login page, an account and password login page, etc., which can be determined based on the actual application scenario requirements and is not limited here.

[0044] Optionally, in response to a user's login selection instruction, the above login page may be displayed to the user through the target terminal, so that the user logs in to the terminal system of the target terminal based on the login page.

[0045] Among them, the user can log in to the terminal system of the target terminal by scanning the QR code on the QR code login page through the designated application, or can log in to the terminal system of the target terminal by entering the corresponding account and password. After logging in to the terminal system of the target terminal, the user has the right to use the target terminal, and can customize the tone for the user based on the audio data uploaded by the user or related setting instructions.

[0046] like Figure 2 As shown, Figure 2 : This is a scene diagram of the login page of the terminal system provided in an embodiment of the present application. The target terminal can display a QR code login page to the user to prompt the user to scan the QR code in the login page through a specified application to log in. Further, the login information sent by the user when scanning the QR code can be obtained, and the user's login status can be verified based on the login information sent by the user, so as to determine whether the user has successfully logged in. If it is determined that the user has failed to log in to the terminal system of the target terminal, a failure prompt message can be displayed to the user through the above QR code login page to instruct the user to log in again.

[0047] Optionally, if it is determined that the user terminal corresponding to the user establishes a communication connection with the target terminal, the terminal system of the target terminal logged in by the user can be determined. If it is determined that the user terminal corresponding to the user establishes a Bluetooth connection with the target terminal, or it is determined that the user terminal corresponding to the user and the target terminal are connected to the same local area network, the terminal system of the target terminal logged in by the user can be determined.

[0048] Furthermore, in response to the user logging into the terminal system of the target terminal, a timbre customization prompt page is displayed to the user. Specifically, the timbre customization prompt page can be displayed to the user through the target terminal after the user logs into the terminal system of the target terminal, or the timbre customization prompt page can be displayed through the user terminal of the user after the user logs into the terminal system of the target terminal, so as to guide the user to customize the timbre of the target terminal during voice broadcasting through the timbre customization prompt page.

[0049] Among them, when displaying the tone customization prompt page, different target terminals may correspond to different tone customization prompt pages. For example, for the target terminal, different tone customization prompt pages are displayed according to one or more of the target terminal model, channel number, and application type (mobile terminal, vehicle-mounted terminal, etc.), which can be determined based on the actual application scenario requirements and are not limited here.

[0050] Among them, when the tone customization prompt page is displayed through the target terminal, the tone customization interface provided by the target terminal through the SDK (Software Development Kit) can be called, and the page configuration information of the tone customization prompt page corresponding to the target terminal is obtained based on the tone customization interface, and then the corresponding tone customization prompt page is displayed to the user through the target terminal based on the page configuration information.

[0051] In some feasible implementations, for the target terminal, the application type, manufacturer, model and channel number of the target terminal will also affect whether the target terminal can customize the tone. For example, the terminal of a certain channel number does not have the configuration conditions for tone customization, a certain type of terminal does not belong to the range of tone customization, the terminal of a certain manufacturer is outside the tone customization whitelist, etc. The specific requirements can be determined based on the actual application scenario and are not restricted here.

[0052] Therefore, when the timbre customization prompt page is displayed through the target terminal, the terminal information of the target terminal including the above-mentioned information can also be determined, and whether the target terminal has the timbre customization authority can be determined based on the terminal information of the target terminal. If the target terminal has the timbre customization authority, the timbre customization prompt page can be further displayed through the target terminal based on the above-mentioned implementation method.

[0053] In the case where the target terminal does not have the timbre customization authority, when the target terminal needs to broadcast audio information by voice, the target terminal can play the audio information with the default timbre corresponding to the default timbre configuration information based on the default timbre configuration information corresponding to the target terminal. The default timbre configuration information corresponding to the target terminal can be stored in a server corresponding to the target terminal, or in a local storage space of the target terminal, which is not limited here.

[0054] Among them, when audio information is played to the user through the target terminal with the default tone corresponding to the default tone configuration information, the application corresponding to the audio information to be played by the target terminal can also be determined, and the default tone corresponding to the application is determined based on the default tone data, and then the audio information to be played corresponding to the application is played through the target terminal with the default tone corresponding to the application.

[0055] Step S12: Acquire the first audio data uploaded by the user based on the timbre customization prompt page, and display the timbre list page.

[0056] In some feasible implementations, the tone customization prompt page includes a tone recording control, which can respond to a confirmation instruction triggered based on the tone recording control and display the tone recording page to obtain the first audio data uploaded by the user through the tone recording page.

[0057] The first audio data uploaded by the user may be audio data recorded by the user based on the timbre recording page, or may be audio data obtained by the user based on local storage space or through the network, which may be determined based on the actual application scenario requirements and is not limited here. In addition, the first audio data uploaded by the user may be audio data with a duration exceeding the duration threshold, or may be multiple audio data corresponding to the same timbre, which is not limited here.

[0058] As an example, a user may upload the broadcast voice of a news anchor as the first audio data, and thus the timbre may be determined as the timbre configuration information of the news anchor based on the first audio data.

[0059] Specifically, in response to a confirmation instruction triggered by a user through a timbre customization prompt page, such as a confirmation instruction triggered by a timbre recording control in the timbre customization prompt page, a login page of the timbre recording program may be displayed. Further, in response to a user logging into the timbre recording program, a timbre recording page may be displayed to guide the user to perform timbre customization after logging into the timbre recording program.

[0060] like Figure 3 As shown, Figure 3 1 is a scene diagram of the voice recording program login page provided in an embodiment of the present application. The QR code login page of the voice recording program is displayed to the user through the target terminal or the user terminal, and the user can log in after recognizing the QR code based on the specified application. Further, based on the user's login information, it can be determined whether the user has logged in to the voice recording program. After the user successfully logs in to the voice recording program, the voice recording page can be displayed through the target terminal or the user terminal, and then the first audio data uploaded by the user can be obtained through the voice recording page.

[0061] Optionally, the above-mentioned timbre recording page includes recording text information, that is, the recording text information can be displayed through the timbre recording page, and the recording text information is used to prompt the user to record audio data based on the recording text information, that is, the user needs to read the recording text information aloud to record the first audio data.

[0062] Based on this, when the text content corresponding to the audio data recorded by the user is consistent with the recorded text information, the first audio data can be obtained. When the text content corresponding to the first audio data recorded by the user is inconsistent with the recorded text information, an error prompt message is displayed through the timbre recording page to prompt the user that the first audio data recorded by the user is different from the recorded text information and the first audio data needs to be re-recorded based on the recorded text information.

[0063] Among them, when the user records the first audio data based on the recorded text information, the audio data recording progress and relevant prompt information for reminding the user that recording are in progress can be displayed in real time through the tone recording page based on the audio data recorded by the user. The specific information can be determined based on the actual application scenario requirements and is not limited here.

[0064] like Figure 4 As shown, Figure 4 It is a scene diagram of the tone recording page provided in an embodiment of the present application. Figure 4 The timbre recording page shown displays the recording text information "The clever ones work hard while the knowledgeable ones worry, and the incompetent ones have nothing to ask for. When they are well fed, they travel freely, floating like an untied boat". The user needs to read the recording text information while long pressing the voice input control to record the first audio data. That is, the text content of the first audio data recorded by the user needs to be consistent with the recording text information displayed in the timbre recording page in order to successfully upload the recorded first audio data. In this process, the timbre recording page can display the prompt information "Recording" to prompt the user that the audio data is currently being recorded. Based on the text content of the audio data recorded by the user and the recording text information, the user's audio recording progress can be determined, and the audio recording progress can be displayed through the timbre recording page.

[0065] In some feasible implementations, after determining that the user logs into the tone recording program, a guide page may be further displayed to guide the user to upload audio data through the guide page.

[0066] As an example, after determining that the user logs into the sound recording program, the first guide page of the sound recording program may be displayed to determine the upload time of the audio data uploaded by the user through the first guide page.

[0067] like Figure 5a As shown, Figure 5a Schematic diagram of the first guide page provided in the embodiment of the present application. Figure 5a The first guide page shown can display relevant descriptive information about the customized timbre to the user, and provide the user with a time option for uploading audio data. For example, two time options, "Customize later" and "Customize now", can be displayed to the user. If the user selects the "Customize now" option, the timbre recording page can be displayed to the user, and the first audio data uploaded by the user can be obtained through the timbre recording page. The above time options can be specifically determined based on the actual application scenario requirements and are not limited here.

[0068] If the user selects the "customize later" option, the timbre recording page is further displayed to the user at a corresponding time based on the user's customization time requirement to obtain the first audio data uploaded by the user based on the timbre recording page.

[0069] It should be noted that the above-mentioned audio data upload time is only an example, which can be determined based on user settings and actual application scenario requirements and is not limited here.

[0070] As an example, after determining that the user has logged into the sound recording program, or after determining based on the first guide page that the user currently needs to record audio data, a second guide page can be displayed to prompt the user whether there is any uploaded audio data, and guide the user to upload the audio data based on relevant controls.

[0071] like Figure 5b As shown, Figure 5b Schematic diagram of the second guide page provided in the embodiment of the present application. Figure 5b The second guide page shown may prompt the user that no audio data has been uploaded before, and provide the user with a "customize exclusive timbre" control to guide the user to upload audio data. If a confirmation instruction triggered by the user based on the "customize exclusive timbre" control is obtained based on the guide page, the timbre recording page may be displayed to the user, and the first audio data uploaded by the user based on the timbre recording page may be obtained.

[0072] As an example, after determining that the user has logged into the timbre recording program, or after determining that the audio data needs to be uploaded based on any of the above-mentioned guide pages, the third guide page of the timbre recording program may be displayed to the user. The third guide page may display to the user the timbre type corresponding to the first audio data that the user needs to upload, thereby determining the timbre type corresponding to the first audio data uploaded by the user. Based on the timbre type selected by the user, when determining the target timbre configuration information based on the first audio data uploaded by the user, the accuracy of the timbre corresponding to the target timbre configuration information may be improved.

[0073] like Figure 5c As shown, Figure 5c Schematic diagram of the third guide page provided in the embodiment of the present application. Figure 5c The third guide page shown can provide the user with a variety of timbre type options, such as "female voice", "male voice", "male child voice" and "female child voice". If the user selects the "female voice" type, it can be determined that the timbre corresponding to the first audio data uploaded by the user thereafter is of the "female voice" type, that is, the timbre customized based on the uploaded first audio data that the user expects is of the "female voice" type. Based on this, after determining the timbre type selected by the user, the timbre type corresponding to the first audio data uploaded by the user is confirmed, and the timbre recording page is displayed to obtain the first audio data uploaded by the user based on the timbre recording page.

[0074] Optionally, if the user has finished recording the first audio data based on the recorded text information, an audio information completion page may be displayed. Further, in response to an upload instruction triggered by the user based on the audio information completion page, the first audio data uploaded by the user is obtained. Among them, a prompt message "audio recording completed" may be displayed to the user through the audio information completion page to prompt the user to complete the recording, and the user's modification information on the first audio data, such as the naming information of the first audio data, may also be obtained through the audio information completion page, and then the modification information is obtained together with the upload instruction triggered by the user based on the audio response page, and the improved first audio data is obtained.

[0075] As an example, Figure 6 As shown, Figure 6 Schematic diagram of the scene of the audio information improvement page provided by the embodiment of the present application. Figure 6 The audio information improvement page shown can provide the user with improvement information for the first audio data, thereby obtaining the name information of the first audio data input by the user based on the audio improvement information page. At the same time, the audio information improvement page can display relevant controls for confirming the uploaded audio data, and in response to the upload instruction triggered by the user based on the space, the first audio data uploaded by the user is obtained.

[0076] In some feasible implementations, after obtaining the first audio data uploaded by the user and confirming that the user has completed uploading the first audio data, a timbre list page may be displayed to the user through the target terminal or the user terminal, and the timbre list page includes target timbre configuration information determined by the first audio data. That is, after obtaining the first audio data uploaded by the user, target timbre configuration information corresponding to the first audio data, such as a voice package, may be generated based on the first audio data. Among them, the timbre corresponding to the target timbre configuration information is the same as the timbre corresponding to the first audio data, that is, target timbre configuration information of the same timbre may be generated based on the first audio data uploaded by the user.

[0077] Optionally, the above-mentioned timbre list page can also be based on the timbre configuration information determined by the historical audio data uploaded by the user, and at least one or more of the at least one default timbre configuration information corresponding to the target terminal. Among them, the timbres corresponding to any two timbre configuration information in the above-mentioned timbre list page are two different timbres. Specifically, the timbre configuration information of each timbre corresponding to the user can be obtained from the server corresponding to the target terminal, and the timbre list page can be generated based on each timbre configuration information.

[0078] Optionally, each timbre configuration information can be displayed as a corresponding timbre type in the timbre list page, that is, the timbre type can be displayed in the timbre list page to represent the corresponding timbre configuration information. And the timbre corresponding to each timbre configuration information can be previewed through the timbre list page, which can be determined based on the actual application scenario requirements and is not limited here.

[0079] As an example, Figure 7a As shown, Figure 7a This is a schematic diagram of a scene of a tone list page provided in an embodiment of the present application. Figure 7a The timbre list page shown can display the timbre types and corresponding avatars corresponding to different timbre configuration information to the user, so that the user can intuitively determine the timbre corresponding to each timbre configuration information. In addition, the timbre list page can display the playback controls corresponding to each timbre type, so that after obtaining the user's play or pause instructions for any playback control based on the timbre list page, the corresponding timbre can be previewed or paused.

[0080] Optionally, before displaying the timbre list page, the timbre attribute information of each timbre configuration information corresponding to the user may also be determined. For any timbre configuration information, the timbre attribute information corresponding to the timbre configuration information includes at least one of the data size of the timbre configuration information, the corresponding timbre type, the information identifier of the timbre configuration information, or the storage path. Then, when displaying the timbre list page, the timbre list page may be displayed based on the timbre attribute information of each timbre configuration information corresponding to the user.

[0081] That is, in addition to displaying the timbre types of different timbre configuration information, the timbre list page can also display information such as the data size of each timbre configuration information. The specific information can be determined based on the actual application scenario requirements and is not limited here. Figure 7b , Figure 7b This is another schematic diagram of a timbre list page provided in an embodiment of the present application. If each timbre configuration information is regarded as a voice package with different timbres, based on Figure 7b The timbre list page shown can display the timbre type of each voice package, such as "little fairy" or "little fresh meat", and can also synchronously display relevant information such as the data size of each voice package. And because the timbre list page is determined based on the timbre attribute information of each voice package, if a download instruction for any voice package is detected based on the timbre list page, the voice package (i.e., timbre configuration information) can be sent to the target terminal so that the terminal can store the timbre information locally.

[0082] Step S13: In response to the user's setting instruction for the target timbre configuration information in the timbre list page, the target terminal plays the audio information with the timbre corresponding to the target timbre configuration information.

[0083] In some feasible implementations, after the timbre list page is displayed, in response to a user's setting instruction for the target timbre configuration information in the timbre list page, the target terminal plays the audio information with the timbre corresponding to the target timbre configuration information. The target timbre configuration information may be any timbre configuration information in the timbre list page.

[0084] Specifically, in response to the user's setting instruction for the target timbre configuration information in the timbre list page, the target timbre configuration information is determined to be synchronized to each application of the target terminal, and then the audio information of any application can be played through the target terminal with the timbre corresponding to the target timbre configuration information. That is, based on the user's setting instruction for the target timbre configuration information in the timbre list page, the timbre corresponding to the target timbre configuration information can be determined as the broadcast timbre of all applications in the target terminal. When the target terminal needs to play the audio information of any application, the corresponding audio information can be played based on the timbre corresponding to the target timbre configuration information.

[0085] Optionally, in response to the user's setting instruction for the target timbre configuration information in the timbre list page, the target application corresponding to the target timbre configuration information is determined, and then the audio information of the target application can be played through the target terminal with the timbre corresponding to the target timbre configuration information. That is, based on the user's setting instruction for the target timbre configuration information in the timbre list page, the timbre corresponding to the target timbre configuration information can be determined as the announcement timbre of the target application corresponding to the setting instruction. When the target terminal needs to play the audio information of the target application, the audio information of the target application can be played based on the timbre corresponding to the target timbre configuration information.

[0086] Optionally, the above-mentioned timbre list page includes a timbre setting control, and in response to a touch instruction triggered by the user based on the timbre setting control in the timbre list page, the timbre setting page is displayed. The above-mentioned timbre setting page can be a timbre setting page for any application, and the timbre setting page includes various application scenarios of the application, timbre configuration information of at least one timbre customized by the user, and at least one default timbre configuration information. The timbre configuration information set by the user for any application scenario of the application can be determined based on the user's setting instruction, so as to customize the timbre of the announcement corresponding to the application in different application scenarios.

[0087] Alternatively, the above-mentioned timbre setting page may be a timbre setting page for each application, and the timbre setting page includes at least one application, application scenarios corresponding to each application, and multiple timbre configuration information including user-customized timbre configuration information and default timbre configuration information. Based on the timbre setting page, the timbre of the announcement corresponding to different application scenarios can be customized for one or more users at the same time. Alternatively, the timbre of the announcement corresponding to the same application scenario can be customized, that is, the announcement timbre of different application scenarios in the same application scenario is the same.

[0088] like Figure 8 As shown, Figure 8 It is a scene diagram of the tone setting page provided in an embodiment of the present application. Figure 8 The timbre setting page for the navigation application shown includes different voice broadcasting scenarios in the navigation application, such as electronic newspapers, road conditions ahead, and safety tips. At this time, in response to the user's instructions to set any timbre configuration information for any application scenario, such as in response to the user's instructions to set a "standard male voice" for the road conditions ahead application scenario, the timbre of the audio information of the target terminal in the road conditions ahead application scenario of the navigation application is determined to be "standard male voice".

[0089] In some feasible implementations, when obtaining the setting instruction of the user for the target timbre configuration information, it can be obtained based on an instant messaging protocol, such as the setting instruction of the user for the target timbre configuration information can be obtained based on a message queuing telemetry transport (MQTT) protocol.

[0090] In some feasible implementations, when playing audio information with the timbre corresponding to the target timbre configuration information through the target terminal, the audio information to be played can be determined first, and the audio information to be played can be processed based on the target timbre configuration information to obtain audio information with the timbre corresponding to the target timbre configuration information, and then the processed audio information can be played through the target terminal.

[0091] Among them, the above-mentioned audio information to be played can be any information that needs to be output to the user in voice through the target terminal by any application, including but not limited to navigation information, human-computer interaction information, etc., which can be determined based on the actual application scenario requirements and is not limited here. For example, the navigation application of the target terminal needs to broadcast driving instructions through the target terminal during the navigation process, such as the translation application of the target terminal determines the translation text corresponding to the text to be translated input by the user and needs to be voice broadcast to the user through the target terminal, etc. That is, the above-mentioned audio information to be played can be information generated by the application in the target terminal itself and needs to be output to the user, or it can be the corresponding result information determined by the application in the target terminal based on the user's query information.

[0092] Specifically, the target terminal can obtain query information input by the user. The query information can be voice information input by the user to each application through the target terminal, or text information input by the user to each application through the target terminal. The specific information can be determined based on the actual application scenario requirements and is not limited here.

[0093] Furthermore, after obtaining the query information input by the user, the text content corresponding to the query information can be determined, and the result information of the query information can be determined based on the text content corresponding to the query information. Specifically, the query information can be speech recognized based on speech technology to obtain the corresponding text information. Optionally, the query information can be text analyzed by natural language processing (NLP) technology to obtain the corresponding text content and semantics. And the result information corresponding to the query information can be obtained by querying the text content corresponding to the result information, predicting through a prediction model, etc.

[0094] The above prediction models include but are not limited to translation models, dialogue models, etc., which can be constructed based on machine learning and other methods. Machine Learning (ML) is a multi-disciplinary cross-disciplinary subject involving probability theory, statistics, approximation theory, convex analysis, algorithm complexity theory and other disciplines. Based on machine learning and deep learning, machines can simulate or realize human learning behavior to acquire new knowledge or skills, reorganize existing knowledge structures to continuously improve their performance, and thus obtain corresponding translation models and dialogue models.

[0095] In some feasible implementations, each timbre configuration information can be stored in a server, a database, a cloud storage or a blockchain. When a target terminal needs to play audio information, the corresponding timbre configuration information can be obtained online and based on the corresponding timbre configuration information, the audio information can be played through the target terminal with the timbre corresponding to the timbre configuration information.

[0096] Among them, the database can be simply regarded as an electronic file cabinet - a place to store electronic files, which can be used to store various timbre configuration information in this application. Blockchain is a new application mode of computer technologies such as distributed data storage, point-to-point transmission, consensus mechanism, encryption algorithm, etc. Blockchain is essentially a decentralized database, a string of data blocks generated by cryptographic methods. In this application, each data block in the blockchain can store various timbre configuration information. Among them, cloud storage is a new concept extended and developed from the concept of cloud computing. It refers to the use of cluster applications, grid technology, and distributed storage file systems. The functions of a large number of different types of storage devices (storage devices are also called storage nodes) in the network are brought together through application software or application interfaces to work together to store various timbre configuration information.

[0097] Optionally, in response to a timbre download instruction triggered by a user based on a timbre list page or a timbre configuration page, the corresponding timbre configuration information can be obtained from a server, blockchain, database, etc. corresponding to the target terminal and sent to the target terminal, so that the target terminal stores the corresponding timbre configuration information. Thus, when the target terminal needs to broadcast audio information with the timbre corresponding to the timbre configuration information, the audio information to be broadcast can be converted into voice based on the timbre configuration information stored locally, and audio information with the timbre corresponding to the timbre configuration information can be obtained, and the processed audio information can be played to the user.

[0098] In some feasible implementations, the timbre configuration information used by the target terminal when playing audio information can be associated with the user, that is, after the user logs in to the terminal system of the target terminal, the corresponding audio information can be played through the target terminal with the timbre configuration information set by the user. After the user logs out, the target terminal and the user are no longer associated, and the audio information can be played through the target terminal with the default timbre corresponding to the default timbre configuration information.

[0099] That is, if the user logs out of the terminal system of the target terminal, the timbre configuration information corresponding to the target terminal can be restored to the default timbre configuration information. Then, after the user logs out, based on the default timbre configuration information, the audio information is played through the target terminal with the default timbre corresponding to the default timbre configuration information.

[0100] If each application in the target terminal is set with corresponding timbre configuration information, the timbre configuration information corresponding to each application can be restored to the corresponding default timbre configuration information after the user logs out. For any application in the target terminal, based on the default timbre configuration information corresponding to the application, the audio information of the application can be played through the target terminal with the default timbre corresponding to the default timbre configuration information.

[0101] The following uses the vehicle terminal as an example to further illustrate the voice playback method provided in the embodiment of the present application. Fig. 9 As shown, Fig. 9 It is a flowchart of the vehicle-side timbre customization provided in an embodiment of the present application, and the flowchart can be applied to the TTS component in the vehicle-side. The TTS component can determine whether the user has logged into the vehicle-side system through the timbre setting entrance of the vehicle-side. If it is determined that the user has logged into the vehicle-side system, the timbre customization prompt page can be displayed through the vehicle-side. If it is determined that the user has not logged into the vehicle-side system, the vehicle-side login page can be displayed through the vehicle-side to prompt the user to log in. Further in response to the user logging into the vehicle-side system, the timbre customization prompt page can be displayed through the vehicle-side. Specifically, based on the personalized customization interface provided by the vehicle-side SDK, a timbre customization prompt page can be popped up, and the timbre customization prompt page can be a user page (UserInterface, UI).

[0102] Among them, when determining whether the user is logged in, it can be determined based on information such as the user identification, and the tone customization prompt page displayed after the user completes the login includes relevant controls for prompting the user to customize the tone, such as a touch button for "recording voice package", and at the same time loading the tone configuration information previously downloaded by the user.

[0103] Furthermore, if it is confirmed based on the tone customization prompt page that the user needs to upload audio data, the login page of the tone recording program can be displayed through the vehicle terminal to allow the user to log in to the tone recording program. After the user completes the login, the tone recording page is displayed to obtain the audio data recorded by the user through the tone recording page and upload the audio data to the Internet of Vehicles server.

[0104] Furthermore, a timbre list page may be displayed to the user, which includes default timbre configuration information and timbre configuration information determined based on audio data recorded by the user, etc. Furthermore, if it is determined that the vehicle terminal has the timbre customization authority, each timbre configuration information in the timbre list page may be obtained and stored in a specific data structure, such as data based on a TTSdata structure may be used to save each timbre configuration information.

[0105] When saving each timbre configuration information, the identification information, name, timbre type, data size, storage path and other information of each timbre configuration information may be saved together, and the above timbre list page may be implemented through RecyclerView.Adapter.

[0106] Among them, in response to the user's download instruction for any tone configuration information, the tone configuration information can be sent to the vehicle computer based on the storage path corresponding to the tone configuration information. This can be specifically achieved through a uniform resource locator (URL) connection (Connection) that supports specific functions of the HyperText Transfer Protocol (HTTP).

[0107] After the vehicle terminal completes downloading the timbre configuration information, it can respond to the setting instruction of any timbre configuration information on the vehicle terminal and synchronize the timbre corresponding to the timbre configuration information to each application on the vehicle terminal. When the vehicle terminal needs to play audio information, it can broadcast the timbre corresponding to the timbre configuration information in voice.

[0108] Combine the following Fig.10a The functions of the above TTS components are further explained. Fig.10a It is a functional framework diagram of the TTS component provided in the embodiment of the present application. After the user logs in to the system on the vehicle side, the TTS component can determine whether the vehicle side uses the default playback engine based on the channel number and other information on the vehicle side. If it is determined that the vehicle side uses the default playback engine, the vehicle side broadcasts the audio information with the default tone when performing voice broadcast. If it is determined that the vehicle side uses the TTS component to customize the tone, the user side of the TTS component can determine the tone configuration information according to the audio data recorded by the user by initializing TTS, and synchronize it to each application, confirm the user's login to the tone customization program through the account service interface, and ensure timely upload of the tone configuration information through the network interface. In addition, the service side of the TTS component can provide users with offline TTS services or online TTS services, that is, the TTS component can play audio information with customized tone through the vehicle side based on the locally stored tone configuration information, and can also obtain the tone configuration information in real time through the network, and play audio information with customized tone through the vehicle side based on the acquired tone configuration information. The user side and the service side of the TTS component are connected through a fixed logic language, such as the Android Interface Definition Language (AIDL).

[0109] See also Fig.10b , Fig.10b This is a schematic diagram of the process of using the TTS component provided in the embodiment of the present application. Fig.10bAs shown, the TTS component can provide offline TTS services and online TTS services. When providing offline TTS services, audio information can be played through the vehicle terminal with customized timbre or default timbre. When playing audio information through the vehicle terminal with customized timbre, the best solution adapted to the vehicle terminal can be selected based on the TTS service priority.

[0110] If the vehicle-side has a TTS application package stored in it, the TTS application package will be used first to provide the vehicle-side with tone customization and voice playback services. The above application packages include but are not limited to Android application packages (Android application package, APK) and application packages that support other operating systems. If the TTS application package is not stored, then if the TTS SDK is stored in the vehicle-side, the vehicle-side will be provided with tone customization and voice playback services based on the built-in TTS SDK. Otherwise, the default TTS component on the vehicle-side will be used to provide the vehicle-side with tone customization and voice playback services. If the above TTS services cannot be performed normally, the vehicle-side can be provided with tone customization and voice playback services based on the highest version of the accompanying TTS service, that is, based on the TTS component compatible with the old version.

[0111] See also Fig.11a , Fig.11a Schematic diagram of the timing of TTS service selection provided by the embodiment of the present application. Fig.11a As shown, after obtaining the channel number of the vehicle terminal (or other terminal information), the TTS management module of the TTS component can determine the TTS service corresponding to the vehicle terminal, online TTS service or offline TTS service. If the TTS service corresponding to the vehicle terminal is an offline TTS service, the voice playback engine selected by the vehicle terminal is determined. If the default playback engine of the vehicle terminal is used, the voice broadcast can be made through the vehicle terminal with the default tone. If the playback engine corresponding to the TTS component is used, the voice broadcast can be made based on the tone configuration information set by the user. If the TTS service corresponding to the vehicle terminal is an online TTS service, the voice playback engine selected by the vehicle terminal is determined to be the playback engine corresponding to the TTS component.

[0112] See also Fig.11b , Fig.11b 1 is a timing diagram of setting the tone provided in the embodiment of the present application. Fig.11aAs shown, the TTS management module of the TTS component can respond to the voice customization related instructions of the vehicle terminal and enter the display page, such as displaying the voice customization prompt page and the voice customization page to obtain the audio data recorded by the user. At the same time, the logic module can obtain the voice list corresponding to the user from the Internet of Vehicles server, and display the various voices corresponding to the user (i.e., the voice configuration information in this application) through the voice list page, including the voice recorded based on the audio data recorded by the user. Furthermore, based on the logic module and the TTS service module, the corresponding voice setting can be completed according to the user's relevant setting instructions, such as determining the voice broadcast voice of the vehicle terminal as the "boy voice male" type of voice, and the TTS service module can synchronize this type of voice to the Internet of Vehicles server, so that the Internet of Vehicles server synchronizes the voice to each application on the vehicle terminal, so that each application on the vehicle terminal uses the same voice.

[0113] By using the embodiment of the present application, after the user logs in to the terminal system of the target terminal, the audio data uploaded by the user can be obtained based on the tone customization prompt page, so that the tone can be customized based on the audio data uploaded by the user to improve the user experience. At the same time, based on the user's setting instructions for each tone configuration information in the tone list page, the tone can be quickly and conveniently customized for any application scenario of any application in the target terminal, further improving the user experience.

[0114] See also Fig.12 , Fig.12 : is a structural diagram of a voice playback device provided in an embodiment of the present application. The voice playback device provided in an embodiment of the present application includes:

[0115] The prompt page display module 121 is used to display a tone customization prompt page in response to a user logging into the terminal system of the target terminal;

[0116] The timbre list display module 122 is used to obtain the first audio data uploaded by the user based on the timbre customization prompt page, and display a timbre list page, wherein the timbre list page includes the target timbre configuration information determined by the first audio data, and the first audio data and the target timbre configuration information correspond to the same timbre;

[0117] The voice playing module 123 is used to respond to the setting instruction of the user for the target timbre configuration information in the timbre list page, and play the audio information with the timbre corresponding to the target timbre configuration information through the target terminal.

[0118] In some feasible implementations, the timbre list display module 122 is used to:

[0119] In response to a confirmation instruction triggered by the user through the timbre customization prompt page, displaying a login page of the timbre recording program;

[0120] In response to the user logging into the timbre recording program, displaying a timbre recording page;

[0121] The first audio data uploaded by the user based on the timbre recording page is obtained.

[0122] In some feasible implementations, the timbre recording page includes recording text information, and the recording text information is used to prompt the user to record audio data based on the recording text information;

[0123] The timbre list display module 122 is used to:

[0124] In response to the text content corresponding to the first audio data recorded by the user being consistent with the recorded text information, the first audio data is acquired.

[0125] In some feasible implementations, the voice playback module 123 is further used to:

[0126] Synchronizing the target timbre configuration information to each application of the target terminal;

[0127] In response to the user's setting instruction for the target timbre configuration information in the timbre list page, the target terminal plays the audio information of any of the above applications with the timbre corresponding to the target timbre configuration information.

[0128] In some feasible implementations, the voice playback module 123 is used to:

[0129] In response to the user's setting instruction for the target timbre configuration information in the timbre list, determining a target application corresponding to the target timbre configuration information;

[0130] The audio information of the target application is played through the target terminal with the timbre corresponding to the target timbre configuration information.

[0131] In some feasible implementations, the prompt page display module 121 is further used to:

[0132] Determine the timbre attribute information of each timbre configuration information corresponding to the above user, wherein the timbre attribute information of any of the above timbre configuration information includes at least one of the data size of the timbre configuration information, the corresponding timbre type, the information identifier or the storage path;

[0133] The timbre list page determined by the timbre attribute information described above is displayed.

[0134] In some feasible implementations, the voice playback module 123 is used to:

[0135] Obtaining the query information input by the user, and determining the text content corresponding to the query information;

[0136] The result information corresponding to the query information is determined based on the text content corresponding to the query information, and the result information is played through the target terminal with the timbre corresponding to the target timbre configuration information.

[0137] In some feasible implementations, the prompt page display module 121 is used to:

[0138] In response to determining that the target terminal has the timbre customization authority based on the terminal information of the target terminal, a timbre customization prompt page is displayed.

[0139] In some feasible implementations, the voice playback module 123 is further used to:

[0140] In response to determining based on the terminal information of the target terminal that the target terminal does not have the timbre customization authority, the audio information is played through the target terminal with a default timbre corresponding to the target terminal.

[0141] In some feasible implementations, the voice playback module 123 is further used to:

[0142] In response to the user logging out of the terminal system, the audio information is played through the target terminal with a default tone corresponding to the target terminal.

[0143] In some feasible implementations, the voice playback module 123 is further used to:

[0144] Based on the message queue telemetry transmission protocol, the setting instruction of the user for the target timbre configuration information is obtained.

[0145] In a specific implementation, the above-mentioned voice playback device can execute the above-mentioned functions through its built-in functional modules. Figure 1 For the implementation methods provided in each step, please refer to the implementation methods provided in the above steps for details, which will not be repeated here.

[0146] See also Fig.13 , Fig.13 Schematic diagram of the structure of the electronic device provided in the embodiment of the present application. Fig.13As shown, the electronic device 1000 in this embodiment may include: a processor 1001, a network interface 1004 and a memory 1005. In addition, the above-mentioned electronic device 1000 may also include: a user interface 1003, and at least one communication bus 1002. Among them, the communication bus 1002 is used to realize the connection and communication between these components. Among them, the user interface 1003 may include a display screen (Display), a keyboard (Keyboard), and the user interface 1003 may optionally include a standard wired interface and a wireless interface. The network interface 1004 may optionally include a standard wired interface and a wireless interface (such as a WI-FI interface). The memory 1004 may be a high-speed RAM memory, or it may be a non-volatile memory (non-volatile memory), such as at least one disk storage. The memory 1005 may optionally also be at least one storage device located away from the aforementioned processor 1001. As Fig.13 As shown, the memory 1005 as a computer-readable storage medium may include an operating system, a network communication module, a user interface module, and a device control application program.

[0147] exist Fig.13 In the electronic device 1000 shown, the network interface 1004 can provide a network communication function; the user interface 1003 is mainly used to provide an input interface for the user; and the processor 1001 can be used to call the device control application stored in the memory 1005 to achieve:

[0148] In response to the user logging into the terminal system of the target terminal, displaying a tone customization prompt page;

[0149] Acquire the first audio data uploaded by the user based on the timbre customization prompt page, and display a timbre list page, wherein the timbre list page includes the target timbre configuration information determined by the first audio data, and the first audio data and the target timbre configuration information correspond to the same timbre;

[0150] In response to the user's setting instruction for the target timbre configuration information in the timbre list page, the target terminal plays the audio information with the timbre corresponding to the target timbre configuration information.

[0151] In some feasible implementations, the processor 1001 is configured to:

[0152] In response to a confirmation instruction triggered by the user through the timbre customization prompt page, displaying a login page of the timbre recording program;

[0153] In response to the user logging into the timbre recording program, displaying a timbre recording page;

[0154] The first audio data uploaded by the user based on the timbre recording page is obtained.

[0155] In some feasible implementations, the timbre recording page includes recording text information, and the recording text information is used to prompt the user to record audio data based on the recording text information;

[0156] The processor 1001 is used for:

[0157] In response to the text content corresponding to the first audio data recorded by the user being consistent with the recorded text information, the first audio data is acquired.

[0158] In some feasible implementations, the processor 1001 is further configured to:

[0159] Synchronizing the target timbre configuration information to each application of the target terminal;

[0160] In response to the user's setting instruction for the target timbre configuration information in the timbre list page, the target terminal plays the audio information of any of the above applications with the timbre corresponding to the target timbre configuration information.

[0161] In some feasible implementations, the processor 1001 is configured to:

[0162] In response to the user's setting instruction for the target timbre configuration information in the timbre list, determining a target application corresponding to the target timbre configuration information;

[0163] The audio information of the target application is played through the target terminal with the timbre corresponding to the target timbre configuration information.

[0164] In some feasible implementations, the processor 1001 is further configured to:

[0165] Determine the timbre attribute information of each timbre configuration information corresponding to the above user, wherein the timbre attribute information of any of the above timbre configuration information includes at least one of the data size of the timbre configuration information, the corresponding timbre type, the information identifier or the storage path;

[0166] The timbre list page determined by the timbre attribute information described above is displayed.

[0167] In some feasible implementations, the processor 1001 is configured to:

[0168] Obtaining the query information input by the user, and determining the text content corresponding to the query information;

[0169] The result information corresponding to the query information is determined based on the text content corresponding to the query information, and the result information is played through the target terminal with the timbre corresponding to the target timbre configuration information.

[0170] In some feasible implementations, the processor 1001 is configured to:

[0171] In response to determining that the target terminal has the timbre customization authority based on the terminal information of the target terminal, a timbre customization prompt page is displayed.

[0172] In some feasible implementations, the processor 1001 is further configured to:

[0173] In response to determining based on the terminal information of the target terminal that the target terminal does not have the timbre customization authority, the audio information is played through the target terminal with a default timbre corresponding to the target terminal.

[0174] In some feasible implementations, the processor 1001 is further configured to:

[0175] In response to the user logging out of the terminal system, the audio information is played through the target terminal with a default tone corresponding to the target terminal.

[0176] In some feasible implementations, the processor 1001 is further configured to:

[0177] Based on the message queue telemetry transmission protocol, the setting instruction of the user for the target timbre configuration information is obtained.

[0178] It should be understood that in some feasible implementations, the processor 1001 may be a central processing unit (CPU), and the processor may also be other general-purpose processors, digital signal processors (DSP), application specific integrated circuits (ASIC), field-programmable gate arrays (FPGA) or other programmable logic devices, discrete gates or transistor logic devices, discrete hardware components, etc. A general-purpose processor may be a microprocessor or the processor may also be any conventional processor, etc. The memory may include a read-only memory and a random access memory, and provide instructions and data to the processor. A portion of the memory may also include a non-volatile random access memory. For example, the memory may also store information about the type of device.

[0179] In a specific implementation, the electronic device 1000 can execute the above-mentioned functions through its built-in functional modules. Figure 1 For the implementation methods provided in each step, please refer to the implementation methods provided in the above steps for details, which will not be repeated here.

[0180] By using the embodiment of the present application, after the user logs in to the terminal system of the target terminal, the audio data uploaded by the user can be obtained based on the tone customization prompt page, so that the tone can be customized based on the audio data uploaded by the user to improve the user experience. At the same time, based on the user's setting instructions for each tone configuration information in the tone list page, the tone can be quickly and conveniently customized for any application scenario of any application in the target terminal, further improving the user experience.

[0181] The present application also provides a computer-readable storage medium, which stores a computer program and is executed by a processor to implement Figure 1 For the methods provided in each step, please refer to the implementation methods provided in the above steps for details, which will not be repeated here.

[0182] The computer-readable storage medium may be an internal storage unit of the aforementioned voice playback device or electronic device, such as a hard disk or memory of the electronic device. The computer-readable storage medium may also be an external storage device of the electronic device, such as a plug-in hard disk, a smart media card (SMC), a secure digital (SD) card, a flash card, etc. equipped on the electronic device. The computer-readable storage medium may also include a magnetic disk, an optical disk, a read-only memory (ROM) or a random access memory (RAM), etc. Further, the computer-readable storage medium may also include both an internal storage unit of the electronic device and an external storage device. The computer-readable storage medium is used to store the computer program and other programs and data required by the electronic device. The computer-readable storage medium may also be used to temporarily store data that has been output or is to be output.

[0183] The present application embodiment provides a computer program product, which includes a computer program or a computer instruction. When the computer program or the computer instruction is executed by a processor, the voice playback method provided in the present application embodiment is executed. Figure 1 The methods provided in each step.

[0184] The terms "first", "second", etc. in the claims, specification and drawings of the present application are used to distinguish different objects, rather than to describe a specific order. In addition, the terms "including" and "having" and any of their variations are intended to cover non-exclusive inclusions. For example, a process, method, system, product or electronic device that includes a series of steps or units is not limited to the listed steps or units, but optionally includes steps or units that are not listed, or optionally includes other steps or units inherent to these processes, methods, products or electronic devices. Mentioning "embodiment" in this article means that the specific features, structures or characteristics described in conjunction with the embodiment may be included in at least one embodiment of the present application. Displaying the phrase at various locations in the specification does not necessarily refer to the same embodiment, nor is it an independent or alternative embodiment that is mutually exclusive with other embodiments. It is explicitly and implicitly understood by those skilled in the art that the embodiments described herein can be combined with other embodiments. The term "and / or" used in the present application specification and the attached claims refers to any combination of one or more of the items listed in association and all possible combinations, and includes these combinations.

[0185] Those skilled in the art will appreciate that the units and algorithm steps of each example described in conjunction with the embodiments disclosed herein can be implemented by electronic hardware, computer software, or a combination of the two. In order to clearly illustrate the interchangeability of hardware and software, the composition and steps of each example have been generally described in terms of function in the above description. Professional and technical personnel may use different methods to implement the described functions for each specific application, but such implementation should not be considered beyond the scope of this application.

[0186] The above disclosure is only the preferred embodiment of the present application, which cannot be used to limit the scope of rights of the present application. Therefore, equivalent changes made according to the claims of the present application are still within the scope covered by the present application.

Claims

1. A voice playing method, characterized in that: The method comprises: In response to the user logging into the terminal system of the target terminal, displaying a tone customization prompt page; Acquire the first audio data uploaded by the user based on the timbre customization prompt page, and display a timbre list page, wherein the timbre list page includes target timbre configuration information determined by the first audio data, and the first audio data and the target timbre configuration information correspond to the same timbre; Synchronizing the target timbre configuration information to each application of the target terminal; In response to a setting instruction of the user for the target timbre configuration information in the timbre list page, playing audio information with the timbre corresponding to the target timbre configuration information through the target terminal; The step of responding to the user's setting instruction for the target timbre configuration information in the timbre list page, playing the audio information with the timbre corresponding to the target timbre configuration information through the target terminal, comprises: In response to the user's setting instruction for the target timbre configuration information in the timbre list page, the target terminal plays the audio information of any of the applications with the timbre corresponding to the target timbre configuration information.

2. The method according to claim 1, characterized in that The obtaining of the first audio data uploaded by the user based on the tone customization prompt page includes: In response to a confirmation instruction triggered by the user through the timbre customization prompt page, displaying a login page of a timbre recording program; In response to the user logging into the timbre recording program, displaying a timbre recording page; Acquire the first audio data uploaded by the user based on the timbre recording page.

3. The method according to claim 2, characterized in that The timbre recording page includes recording text information, and the recording text information is used to prompt the user to record audio data based on the recording text information; The obtaining the first audio data uploaded by the user based on the timbre recording page includes: In response to the text content corresponding to the first audio data recorded by the user being consistent with the recorded text information, the first audio data is acquired.

4. The method according to claim 1, characterized in that: The step of responding to the user's setting instruction for the target timbre configuration information in the timbre list page, playing the audio information with the timbre corresponding to the target timbre configuration information through the target terminal, comprises: In response to a setting instruction of the user for target timbre configuration information in the timbre list, determining a target application corresponding to the target timbre configuration information; The target terminal plays the audio information of the target application program with the timbre corresponding to the target timbre configuration information.

5. The method according to claim 1, characterized in that The method further comprises: Determine the timbre attribute information of each timbre configuration information corresponding to the user, wherein the timbre attribute information of any timbre configuration information includes at least one of the data size of the timbre configuration information, the corresponding timbre type, the information identifier or the storage path; The page for displaying the tone list includes: A timbre list page determined by the timbre attribute information is displayed.

6. The method according to claim 1, characterized in that The step of playing the audio information with the timbre corresponding to the target timbre configuration information by the target terminal includes: Acquire the query information input by the user, and determine the text content corresponding to the query information; Result information corresponding to the query information is determined based on text content corresponding to the query information, and the result information is played through the target terminal with the timbre corresponding to the target timbre configuration information.

7. The method according to claim 1, characterized in that The display tone customization prompt page includes: In response to determining that the target terminal has the timbre customization authority based on the terminal information of the target terminal, a timbre customization prompt page is displayed.

8. The method according to claim 7, characterized in that The method further comprises: In response to determining that the target terminal does not have the timbre customization authority based on the terminal information of the target terminal, audio information is played through the target terminal with a default timbre corresponding to the target terminal.

9. The method according to claim 1, characterized in that: The method further comprises: In response to the user logging out of the terminal system, the audio information is played through the target terminal with a default tone corresponding to the target terminal.

10. The method according to claim 1, characterized in that The method further comprises: The setting instruction of the user for the target timbre configuration information is obtained based on a message queue telemetry transmission protocol.

11. A voice playing device, characterized in that: The device comprises: A prompt page display module, used for displaying a tone customization prompt page in response to a user logging into a terminal system of a target terminal; A timbre list display module, used to obtain the first audio data uploaded by the user based on the timbre customization prompt page, and display a timbre list page, wherein the timbre list page includes target timbre configuration information determined by the first audio data, and the first audio data and the target timbre configuration information correspond to the same timbre; A voice playing module, used for synchronizing the target timbre configuration information to each application of the target terminal; in response to the user's setting instruction for the target timbre configuration information in the timbre list page, playing audio information with the timbre corresponding to the target timbre configuration information through the target terminal; The voice playback module responds to the user's setting instructions for the target timbre configuration information in the timbre list page, and plays audio information through the target terminal with the timbre corresponding to the target timbre configuration information. It is used to respond to the user's setting instructions for the target timbre configuration information in the timbre list page, and play audio information of any of the applications through the target terminal with the timbre corresponding to the target timbre configuration information.

12. The device according to claim 11, characterized in that When the timbre list display module obtains the first audio data uploaded by the user based on the timbre customization prompt page, it is used to: In response to a confirmation instruction triggered by the user through the timbre customization prompt page, displaying a login page of a timbre recording program; In response to the user logging into the timbre recording program, displaying a timbre recording page; Acquire the first audio data uploaded by the user based on the timbre recording page.

13. The device according to claim 12, characterized in that The timbre recording page includes recording text information, and the recording text information is used to prompt the user to record audio data based on the recording text information; When the timbre list display module obtains the first audio data uploaded by the user based on the timbre recording page, it is used to: In response to the text content corresponding to the first audio data recorded by the user being consistent with the recorded text information, the first audio data is acquired.

14. The device according to claim 11, characterized in that The voice playing module responds to the setting instruction of the user for the target timbre configuration information in the timbre list page, and plays the audio information with the timbre corresponding to the target timbre configuration information through the target terminal, for: In response to a setting instruction of the user for target timbre configuration information in the timbre list, determining a target application corresponding to the target timbre configuration information; The target terminal plays the audio information of the target application program with the timbre corresponding to the target timbre configuration information.

15. The device according to claim 11, characterized in that The prompt page display module is also used for: Determine the timbre attribute information of each timbre configuration information corresponding to the user, wherein the timbre attribute information of any timbre configuration information includes at least one of the data size of the timbre configuration information, the corresponding timbre type, the information identifier or the storage path; A timbre list page determined by the timbre attribute information is displayed.

16. The device according to claim 11, characterized in that When the voice playing module plays the audio information with the timbre corresponding to the target timbre configuration information through the target terminal, it is used to: Acquire the query information input by the user, and determine the text content corresponding to the query information; Result information corresponding to the query information is determined based on text content corresponding to the query information, and the result information is played through the target terminal with the timbre corresponding to the target timbre configuration information.

17. The device according to claim 11, characterized in that When the prompt page display module displays the tone customization prompt page, it is used to: In response to determining that the target terminal has the timbre customization authority based on the terminal information of the target terminal, a timbre customization prompt page is displayed.

18. The device according to claim 17, characterized in that The voice playback module is also used for: In response to determining that the target terminal does not have the timbre customization authority based on the terminal information of the target terminal, audio information is played through the target terminal with a default timbre corresponding to the target terminal.

19. The device according to claim 11, characterized in that The voice playback module is also used for: In response to the user logging out of the terminal system, the audio information is played through the target terminal with a default tone corresponding to the target terminal.

20. The device according to claim 11, characterized in that The voice playback module is also used for: The setting instruction of the user for the target timbre configuration information is obtained based on a message queue telemetry transmission protocol.

21. An electronic device, characterized in that: comprising a processor and a memory, wherein the processor and the memory are connected to each other; The memory is used to store computer programs; The processor is configured to execute the method according to any one of claims 1 to 10 when calling the computer program.

22. A computer-readable storage medium, characterized in that: The computer-readable storage medium stores a computer program, and the computer program is executed by a processor to implement the method according to any one of claims 1 to 10.

23. A computer program product, characterized in that The computer program product comprises a computer program or computer instructions, and when the computer program or the computer instructions are executed by a processor, the method according to any one of claims 1 to 10 is implemented.

Citation Information

Patent Citations

  • Speech synthesis method and device, electronic equipment and storage medium

    CN113160791A