A TTS voice announcement method, device, vehicle, and computer storage medium
By presetting the broadcast information in the existing TTS voice broadcast technology, merging or prioritizing broadcasting, the problem of interruption of broadcasting, too long time or unclear mixing is solved, and the user experience is improved.
Patent Information
- Application Number
- CN202110963722.7
- Authority / Receiving Office
- CN · China
- Patent Type
- Patents(China)
- Current Assignee / Owner
- Filing Date
- 2021-08-20
- Publication Date
- 2025-07-01
- Estimated Expiration
- 2041-08-20
AI Technical Summary
When existing TTS voice broadcasting technology processes multiple broadcast information, it is easy to cause the broadcast to be suddenly interrupted, the broadcast time is too long or the mixing is unclear, resulting in poor user experience.
By obtaining the current TTS broadcast information and the new TTS broadcast information, preset processing is performed to merge or priority broadcasting to ensure that the broadcast is not interrupted and the information is complete.
It realizes that users can obtain complete and streamlined TTS broadcast information without interrupting or mixing, improving user experience.
Smart Images

Figure CN115708154B_ABST
Abstract
Description
Technical Field
[0001] The present invention relates to the field of TTS voice broadcast, and particularly to a TTS voice broadcast method, device, vehicle and computer storage medium. Background Art
[0002] With the development and popularization of computer technology, intelligent technologies such as human-computer interaction provide convenient and fast services in all aspects of people's lives. TTS (Text to Speech) can realize the conversion from text to speech and is an important technology for human-computer interaction in artificial intelligence technology. TTS is widely used in various intelligent terminals, such as in-vehicle intelligent terminals. When the existing TTS performs a broadcast, it usually directly performs a broadcast after receiving a new broadcast message; or, when receiving a new broadcast message, it interrupts the previous broadcast and directly enters the new voice broadcast; or, when receiving multiple broadcast messages, it queues the multiple broadcast messages and broadcasts them in sequence; or, when receiving multiple broadcast messages, it performs a mixed broadcast on the multiple broadcast messages. However, in actual applications, the above processing methods cause the TTS broadcast to be suddenly interrupted, or the broadcast time is too long, or the mixed broadcast is not clear, so that multiple applications using TTS to broadcast simultaneously need to preempt the audio focus, resulting in a poor user experience. Summary of the Invention
[0003] The purpose of the present invention is to provide a TTS voice broadcast method, device, vehicle and computer storage medium, which can enable users to obtain complete and concise TTS broadcast information without interruption or mixing, and improve the user experience.
[0004] To achieve the above object, the technical solution of the present invention is implemented as follows:
[0005] In a first aspect, an embodiment of the present invention provides a TTS voice broadcast method, and the TTS voice broadcast method includes the following steps:
[0006] Obtain the current TTS broadcast information;
[0007] Obtain new TTS broadcast information;
[0008] According to the new TTS broadcast information, perform preset processing on the current TTS broadcast information;
[0009] Perform voice broadcast on the current TTS broadcast information after the preset processing.
[0010] Second aspect, an embodiment of the present invention provides a TTS voice broadcast device, including a memory, a processor, and a computer program stored in the memory and executable on the processor. When the processor executes the computer program, the steps of the TTS voice broadcast method described in the first aspect are implemented.
[0011] Third aspect, an embodiment of the present invention provides a vehicle, and the vehicle includes the TTS voice broadcast device described in the second aspect.
[0012] Fourth aspect, an embodiment of the present invention provides a computer storage medium, in which a computer program is stored. When the computer program is executed by a processor, the steps of the TTS voice broadcast method described in the first aspect are implemented.
[0013] The TTS voice broadcast method, device, vehicle, and computer storage medium provided by the embodiments of the present invention. The TTS voice broadcast method includes the following steps: obtaining current TTS broadcast information; obtaining new TTS broadcast information; performing preset processing on the current TTS broadcast information according to the new TTS broadcast information; and performing voice broadcast on the current TTS broadcast information after the preset processing. In this way, by obtaining the current TTS broadcast information and the new TTS broadcast information, then performing preset processing on the current TTS broadcast information according to the new TTS broadcast information, and performing voice broadcast on the current TTS broadcast information after the preset processing, the user can obtain complete and concise TTS broadcast information without interruption or mixing, improving the user experience. BRIEF DESCRIPTION OF THE DRAWINGS
[0014] Figure 1 It is a schematic flowchart of a TTS voice broadcast method provided by an embodiment of the present invention;
[0015] Figure 2 It is a schematic structural diagram of a TTS voice broadcast device provided by an embodiment of the present invention. DETAILED DESCRIPTION OF THE EMBODIMENTS
[0016] It should be noted that in this document, the terms "include", "comprise" or any other variants thereof are intended to cover non-exclusive inclusion, such that a process, method, article or apparatus comprising a series of elements not only includes those elements but also other elements not expressly listed, or elements inherent to such process, method, article or apparatus. Without further limitation, an element defined by the statement "comprising one..." does not exclude the presence of additional identical elements in the process, method, article or apparatus comprising such element. In addition, components, features, and elements with the same name in different embodiments of the present invention may have the same meaning or different meanings, and their specific meanings need to be determined based on their explanations in the specific embodiments or further in combination with the context in the specific embodiments.
[0017] It should be understood that although the terms first, second, third, etc. may be used herein to describe various information, such information should not be limited to these terms. These terms are only used to distinguish information of the same type from each other. For example, without departing from the scope of this document, the first information may also be referred to as the second information, and similarly, the second information may also be referred to as the first information. Depending on the context, the word "if" as used herein can be interpreted as "when" or "upon" or "in response to determining". Furthermore, as used herein, the singular forms "a", "an" and "the" are also intended to include the plural forms unless the context indicates otherwise. It should be further understood that the terms "comprise", "include" indicate the presence of the stated features, steps, operations, elements, components, items, kinds, and / or groups, but do not exclude the presence, occurrence or addition of one or more other features, steps, operations, elements, components, items, kinds, and / or groups. The term "or" and "and / or" used herein are interpreted as inclusive, or meaning any one or any combination. Thus, "A, B or C" or "A, B and / or C" means "any of the following: A; B; C; A and B; A and C; B and C; A, B and C". An exception to this definition only occurs when the combination of elements, functions, steps or operations is inherently mutually exclusive in some way.
[0018] It should be understood that although the steps in the flowchart in the embodiments of the present invention are displayed in sequence according to the indication of the arrows, these steps are not necessarily executed in the order indicated by the arrows. Unless there is a clear description in this article, the execution of these steps has no strict order limit and can be executed in other orders. Moreover, at least a part of the steps in the figure may include multiple sub-steps or multiple stages. These sub-steps or stages are not necessarily executed at the same time, but can be executed at different times, and their execution order is not necessarily sequential, but can be executed alternately or in turn with at least a part of other steps or sub-steps or stages of other steps.
[0019] It should be noted that in this article, step codes such as S101 and S102 are adopted. The purpose is to more clearly and briefly express the corresponding content and do not constitute a substantial limitation in terms of order. Those skilled in the art may execute S102 first and then S101 during specific implementation, etc., but these should all be within the protection scope of the present invention.
[0020] It should be understood that the specific embodiments described herein are only used to explain the present invention and are not used to limit the present invention.
[0021] See Figure 1 , which is a TTS voice broadcast method provided by an embodiment of the present invention. This TTS voice broadcast method can be executed by a TTS voice broadcast device provided by an embodiment of the present invention. The TTS voice broadcast device can be implemented in a software and / or hardware manner. In this embodiment, the TTS voice broadcast method is applied to an in-vehicle terminal as an example for illustration. The TTS voice broadcast method includes the following steps:
[0022] Step S101: Obtain the current TTS broadcast information;
[0023] It should be noted that the TTS broadcast information is based on the feedback information or instructions obtained by the in-vehicle terminal from the cloud platform, user terminal, or other device terminals, rather than being fixedly set in the in-vehicle terminal, thereby realizing dynamic broadcast according to the vehicle status information. The in-vehicle terminal connects to the user terminal or other device terminals through Bluetooth technology and obtains feedback information or instructions. This is an optional extended function to enrich and enhance the compatibility of the system. If this function is not used, it will not affect the normal operation of the system.
[0024] Step S102: Obtain new TTS broadcast information;
[0025] Specifically, when the currently obtained TTS broadcast information is being broadcast in step S101, there may be a situation where multiple applications need to broadcast the TTS broadcast information simultaneously. At this time, new TTS broadcast information needs to be obtained, that is, the TTS broadcast information that other applications need to broadcast simultaneously.
[0026] Step S103: Perform preset processing on the currently obtained TTS broadcast information according to the new TTS broadcast information;
[0027] Specifically, when the new TTS broadcast information is obtained, that is, when there is a situation where multiple applications need to broadcast the TTS broadcast information simultaneously, it is necessary to process the currently obtained TTS broadcast information according to the new TTS broadcast information so that both the new TTS broadcast information and the currently obtained TTS broadcast information can be broadcast without being affected.
[0028] Step S104: Perform voice broadcast on the currently obtained TTS broadcast information after the preset processing.
[0029] Specifically, perform voice broadcast on the currently obtained TTS broadcast information after the preset processing in step S103.
[0030] In summary, in the TTS voice broadcast method provided by the above embodiments, by obtaining the currently obtained TTS broadcast information and the new TTS broadcast information, then performing preset processing on the currently obtained TTS broadcast information according to the new TTS broadcast information, and performing voice broadcast on the currently obtained TTS broadcast information after the preset processing, the user can obtain complete and concise TTS broadcast information without interruption or mixing, improving the user experience.
[0031] In one embodiment, the performing preset processing on the currently obtained TTS broadcast information according to the new TTS broadcast information includes the following steps:
[0032] When the currently obtained TTS broadcast information has not started to be broadcast, extract the same characters from the currently obtained TTS broadcast information and the new TTS broadcast information;
[0033] Perform merging processing on the extracted same characters.
[0034] Specifically, when the current TTS broadcast information is in the preparation stage for broadcasting, that is, when the current TTS broadcast information has not started broadcasting, the same characters of the current TTS broadcast information and the new TTS broadcast information are extracted, and the extracted same characters are merged. For example, if the current TTS broadcast information is "Opening the window for you, please be careful of pinching your hand by the window", and the new TTS broadcast information is "Opening the sunroof for you", then the same characters of the current TTS broadcast information and the new TTS broadcast information are extracted as "Opening the window for you", and the TTS broadcast information after merging the extracted same characters is "Opening the sunroof, window for you, please be careful of pinching your hand by the window". In this way, when multiple TTS broadcast information needs to be broadcast simultaneously, the same characters of the TTS broadcast information are extracted and merged, enabling the user to obtain complete and concise TTS broadcast information, reducing the duration of TTS broadcast, and enhancing the user experience.
[0035] In one embodiment, the preset processing of the current TTS broadcast information according to the new TTS broadcast information includes the following steps:
[0036] When the current TTS broadcast information is being broadcast, the new TTS broadcast information is merged to the end of the current TTS broadcast information.
[0037] Specifically, when the current TTS broadcast information is being broadcast, that is, when the current TTS broadcast information has already started broadcasting, the new TTS broadcast information is merged to the end of the current TTS broadcast information. For example, if the currently broadcast TTS broadcast information is "Opening the window for you, please be careful of pinching your hand by the window", and the new TTS broadcast information is "Opening the sunroof for you", then the TTS broadcast information after merging the new TTS broadcast information to the end of the current TTS broadcast information is "Opening the window for you, please be careful of pinching your hand by the window, opening the sunroof for you". In this way, when new TTS broadcast information is obtained while the TTS broadcast information is being broadcast, and the new TTS broadcast information is merged to the end of the current TTS broadcast information, the user can hear the complete TTS broadcast information without interruption or mixing, enhancing the user experience.
[0038] In one embodiment, the preset processing of the current TTS broadcast information according to the new TTS broadcast information includes the following steps:
[0039] When the current TTS broadcast information has not started broadcasting, compare the priority of the new TTS broadcast information with the current TTS broadcast information;
[0040] According to the comparison result, preferentially broadcast the TTS broadcast information with a higher priority.
[0041] Here, when the current TTS broadcast information is in the preparation stage for broadcast, that is, when the current TTS broadcast information has not started to be broadcast, the priorities of the new TTS broadcast information and the current TTS broadcast information are compared. According to the comparison result, the TTS broadcast information with a higher priority is preferentially broadcast. For example, when the current TTS broadcast information is obtained as "It is sunny today, 15°C, with a northeast wind of level 3 - 4, a bit cool. It will rain later today...", and the new TTS broadcast information is obtained as "The current vehicle speed is too fast, and the window cannot be opened", at this time, the priority of the new TTS broadcast information is higher than that of the currently broadcast TTS broadcast information, so the TTS broadcast information with a higher priority is preferentially broadcast as "The current vehicle speed is too fast, and the window cannot be opened. It is sunny today, 15°C, with a northeast wind of level 3 - 4, a bit cool. It will rain later today...". In this way, when multiple TTS broadcast information needs to be broadcast simultaneously, the priorities of the TTS broadcast information are compared, and the TTS broadcast information with a higher priority is preferentially broadcast, enabling users to obtain more important information first and further improving the user experience.
[0042] In one embodiment, when the current TTS broadcast information is being prepared for broadcast, comparing the priorities of the new TTS broadcast information and the current TTS broadcast information further includes the following steps:
[0043] The priorities are arranged from high to low as safety reminder, navigation broadcast, vehicle control reminder, multimedia, and other types.
[0044] Here, vehicle control reminders include user - initiated vehicle control, system automatic adjustment (such as air - conditioner temperature, etc.), vehicle - control - type inquiries, etc.; multimedia includes music, radio, video, and each of them includes user - initiated control and prompt - type broadcasts, etc. The priorities are arranged from high to low as safety reminder, navigation broadcast, vehicle control reminder, multimedia, and other types.
[0045] In one embodiment, performing preset processing on the current TTS broadcast information according to the new TTS broadcast information includes the following steps:
[0046] When the current TTS broadcast information is being broadcast, compare the priorities of the new TTS broadcast information and the current TTS broadcast information;
[0047] If the priority of the new TTS broadcast information is higher than that of the current TTS broadcast information, truncate the current TTS broadcast information and insert and broadcast the new TTS broadcast information.
[0048] Specifically, when the current TTS broadcast information is being broadcast, that is, when the current TTS broadcast information has started to be broadcast, the priorities of the new TTS broadcast information and the current TTS broadcast information are compared. Among them, the priorities are arranged from high to low as safety reminder, navigation broadcast, vehicle control reminder, multimedia, and other types. If the priority of the new TTS broadcast information is higher than that of the current TTS broadcast information, the current TTS broadcast information is truncated and the new TTS broadcast information is inserted for broadcast. For example, the current TTS broadcast information being broadcast is "The weather is sunny today, 15°C, with a northeast wind of level 3-4, a bit cool. It will rain later today...", and the new TTS broadcast information obtained is "The current vehicle speed is too fast to support opening the window". At this time, since the priority of the new TTS broadcast information is higher than that of the current TTS broadcast information being broadcast, the current TTS broadcast information is truncated and the new TTS broadcast information is inserted for broadcast as "The weather is sunny today, 15°C, the current vehicle speed is too fast to support opening the window...". In this way, when a higher-priority TTS broadcast information is obtained during the broadcast of the TTS broadcast information, the currently broadcast TTS broadcast information is truncated and the higher-priority TTS broadcast information is inserted for broadcast, enabling the user to obtain more important TTS broadcast information first and further improving the user experience.
[0049] In an embodiment, the step of truncating the current TTS broadcast information and inserting and broadcasting the new TTS broadcast information if the priority of the new TTS broadcast information is higher than that of the current TTS broadcast information further includes the following steps:
[0050] Obtain the punctuation characters of the current TTS broadcast information;
[0051] Truncate at the comma or full stop of the current TTS broadcast information and insert and broadcast the new TTS broadcast information.
[0052] Specifically, when it is necessary to truncate the current TTS broadcast information and insert and broadcast the new TTS broadcast information, it is also necessary to obtain the punctuation characters of the current TTS broadcast information, truncate at the comma or full stop of the current TTS broadcast information, and insert and broadcast the new TTS broadcast information. In this way, truncating at the comma or full stop of the current TTS broadcast information can enable the user to obtain more important TTS broadcast information first while ensuring the semantic integrity of the current TTS broadcast information as much as possible, further improving the user experience.
[0053] Based on the same inventive concept as the foregoing embodiments, an embodiment of the present invention provides a TTS voice broadcast device, as Figure 2 shown. The TTS voice broadcast device includes: a processor 110 and a memory 111 for storing a computer program that can run on the processor 110; wherein,Figure 2 The processor 110 shown in [Figure] does not refer to the number of processors 110 being one, but only refers to the positional relationship of the processor 110 relative to other devices. In practical applications, the number of processors 110 can be one or more; similarly, Figure 2 the same applies to the memory 111 shown in [Figure], that is, it only refers to the positional relationship of the memory 111 relative to other devices. In practical applications, the number of memories 111 can be one or more. When the processor 110 is used to run the computer program, the TTS voice broadcast method is implemented.
[0054] The TTS voice broadcast device may further include: at least one network interface 112. Each component in the TTS voice broadcast device is coupled together through a bus system 113. It can be understood that the bus system 113 is used to realize the connection and communication between these components. In addition to the data bus, the bus system 113 also includes a power bus, a control bus, and a status signal bus. However, for the sake of clear illustration, in Figure 2 all kinds of buses are labeled as the bus system 113.
[0055] It should be noted that there may be some inaccuracies in the translation as the original text seems to be missing some context like "Figure" in the brackets which are assumed for the purpose of translation. You may need to adjust according to the actual complete context.Among them, the memory 111 can be a volatile memory or a non-volatile memory, or can include both volatile and non-volatile memories. Among them, the non-volatile memory can be a read-only memory (ROM, Read Only Memory), a programmable read-only memory (PROM, Programmable Read-Only Memory), an erasable programmable read-only memory (EPROM, Erasable Programmable Read-Only Memory), an electrically erasable programmable read-only memory (EEPROM, Electrically Erasable Programmable Read-Only Memory), a ferromagnetic random access memory (FRAM, ferromagnetic random access memory), a flash memory (Flash Memory), a magnetic surface memory, an optical disc, or a compact disc read-only memory (CD-ROM, Compact Disc Read-Only Memory); the magnetic surface memory can be a disk memory or a tape memory. The volatile memory can be a random access memory (RAM, Random Access Memory), which is used as an external cache. By way of example but not limitation, many forms of RAM are available, such as a static random access memory (SRAM, Static Random Access Memory), a synchronous static random access memory (SSRAM, Synchronous Static Random Access Memory), a dynamic random access memory (DRAM, Dynamic Random Access Memory), a synchronous dynamic random access memory (SDRAM, Synchronous Dynamic Random Access Memory), a double data rate synchronous dynamic random access memory (DDR SDRAM, Double Data Rate Synchronous Dynamic Random Access Memory), an enhanced synchronous dynamic random access memory (ESDRAM, Enhanced Synchronous Dynamic Random Access Memory), a sync link dynamic random access memory (SLDRAM, SyncLink Dynamic Random Access Memory), a direct rambus random access memory (DRRAM, Direct Rambus Random Access Memory).The memory 111 described in the embodiments of the present invention is intended to include, but is not limited to, these and any other suitable types of memories.
[0056] The memory 111 in the embodiments of the present invention is used to store various types of data to support the operation of the TTS voice broadcast device. Examples of such data include: any computer programs for operating on the TTS voice broadcast device, such as operating systems and application programs; contact data; phone book data; messages; pictures; videos, etc. Among them, the operating system contains various system programs, such as the framework layer, the core library layer, the driver layer, etc., for implementing various basic services and processing hardware-based tasks. The application programs can include various application programs, such as a media player and a browser, for implementing various application services. Here, the program for implementing the method of the embodiments of the present invention can be included in the application programs.
[0057] Based on the same inventive concept as the foregoing embodiments, this embodiment further provides a vehicle, which includes the TTS voice broadcast device as described above.
[0058] Based on the same inventive concept as the foregoing embodiments, this embodiment further provides a computer storage medium, in which a computer program is stored. The computer storage medium can be a ferromagnetic random access memory (FRAM), a read-only memory (ROM), a programmable read-only memory (PROM), an erasable programmable read-only memory (EPROM), an electrically erasable programmable read-only memory (EEPROM), a flash memory, a magnetic surface memory, an optical disc, or a compact disc read-only memory (CD-ROM), etc.; it can also be various devices including one or any combination of the above memories, such as a mobile phone, a computer, a tablet device, a personal digital assistant, etc. When the computer program stored in the computer storage medium is run by a processor, the above-described TTS voice broadcast method is implemented. For the specific step flow implemented when the computer program is executed by the processor, please refer to Figure 1 the description of the illustrated embodiments, which will not be elaborated herein.
[0059] The technical features of the above-described embodiments may be combined arbitrarily. For the sake of brevity of description, not all possible combinations of the technical features in the above-described embodiments are described. However, as long as there is no contradiction in the combination of these technical features, it should be considered as falling within the scope described in this specification.
[0060] In this document, the term "comprising", "including" or any other variant thereof is intended to cover non-exclusive inclusion. In addition to the listed elements, it may also include other elements not expressly listed.
[0061] As described above, this is only a specific implementation manner of the present invention, but the protection scope of the present invention is not limited thereto. Any person skilled in the art within the technical scope disclosed by the present invention can easily think of changes or substitutions, which should all be covered within the protection scope of the present invention. Therefore, the protection scope of the present invention should be subject to the protection scope of the claims.
Claims
1. A TTS voice announcement method, characterized in that, the TTS voice announcement method comprises the following steps: Obtain the current TTS announcement information; Obtain new TTS announcement information; According to the new TTS announcement information, perform preset processing on the current TTS announcement information; Perform voice announcement on the preset processed current TTS announcement information; Wherein, the performing preset processing on the current TTS announcement information according to the new TTS announcement information comprises the following steps: When the current TTS announcement information has not started to be announced, extract the same characters from the current TTS announcement information and the new TTS announcement information; Perform merging processing on the extracted same characters.
2. The TTS voice announcement method according to claim 1, characterized in that, the performing preset processing on the current TTS announcement information according to the new TTS announcement information comprises the following steps: When the current TTS announcement information is being announced, fuse the new TTS announcement information to the end of the current TTS announcement information.
3. The TTS voice announcement method according to claim 1, characterized in that, the performing preset processing on the current TTS announcement information according to the new TTS announcement information comprises the following steps: When the current TTS announcement information has not started to be announced, compare the priorities of the new TTS announcement information and the current TTS announcement information; According to the comparison result, preferentially announce the TTS announcement information with a higher priority.
4. The TTS voice announcement method according to claim 3, characterized in that, the comparing the priorities of the new TTS announcement information and the current TTS announcement information when the current TTS announcement information is about to be announced further comprises the following steps: The priorities are arranged from high to low as safety reminder, navigation announcement, vehicle control reminder, multimedia, other types.
5. The TTS voice announcement method according to claim 1, characterized in that, the performing preset processing on the current TTS announcement information according to the new TTS announcement information comprises the following steps: When the current TTS announcement information is being announced, compare the priorities of the new TTS announcement information and the current TTS announcement information; If the priority of the new TTS announcement information is higher than that of the current TTS announcement information, truncate the current TTS announcement information and insert and announce the new TTS announcement information.
6. The TTS voice announcement method according to claim 5, characterized in that, the truncating the current TTS announcement information and inserting and announcing the new TTS announcement information if the priority of the new TTS announcement information is higher than that of the current TTS announcement information further comprises the following steps: Obtain the punctuation characters of the current TTS announcement information; Truncate at the comma or full stop of the current TTS announcement information and insert and announce the new TTS announcement information.
7. A TTS voice broadcast device, comprising a memory, a processor, and a computer program stored in the memory and executable on the processor, characterized in that when the processor executes the computer program, the steps of the TTS voice broadcast method according to any one of claims 1 to 6 are implemented.
8. A vehicle, characterized in that the vehicle includes the TTS voice broadcast device according to claim 7.
9. A computer storage medium storing a computer program, characterized in that when the computer program is executed by a processor, the steps of the TTS voice broadcast method according to any one of claims 1 to 6 are implemented.
Citation Information
Patent Citations
Radio magnetic guidance system
CA441226A
Speech recognition system and speech recognition device
CN110914897A