Audio resource distribution method and related products

By obtaining the structure information of the audio resource, adding additional resources to the audio resource, the problem of cumbersome audio content distribution in the prior art is solved, and efficient audio resource distribution is achieved.

CN116312658BActive Publication Date: 2025-05-09网易有道信息技术(杭州)有限公司
View PDF 2 Cites 0 Cited by

Patent Information

Application Number
CN202310265913.5
Authority / Receiving Office
CN · China
Patent Type
Patents(China)
Current Assignee / Owner
Filing Date
2023-03-14
Publication Date
2025-05-09
Estimated Expiration
2043-03-14

AI Technical Summary

Technical Problem

In the prior art, if new audio content is needed to be added to the audio content when distributing audio content, the source file must be re-recorded, resulting in cumbersome and inefficient.

Method used

By acquiring the structure information of the audio resource, additional resources are added to the audio resource according to the structure information, and the resources to be distributed are formed, and distributed through the determined distribution channel.

Benefits of technology

The source files of audio resources are not required to be re-recorded, which simplifies the distribution process of audio resources and improves distribution efficiency.

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure CN116312658B_ABST
    Figure CN116312658B_ABST
Patent Text Reader

Abstract

The embodiments of the present invention provide a method for distributing audio resources and related products. The method includes: determining the distribution channel and additional resources of the audio resources in response to the distribution demand for the audio resources; obtaining structural information about the audio resources; adding the additional resources to the audio resources according to the structural information of the audio resources to obtain the resources to be distributed; and distributing the resources to be distributed through the distribution channel. Through the technical solution of the present invention, the audio resources can be structured, and the structural information of the audio resources can be used to support the addition of additional resources at any position in the audio resources without re-editing the source files. In this way, the audio resource distribution cycle can be greatly shortened, and the distribution efficiency can be effectively improved.
Need to check novelty before this filing date? Find Prior Art

Description

Technical Field

[0001] Embodiments of the present invention relate to the field of information processing technology, and more specifically, to a method for distributing audio resources, and an electronic device and a computer-readable storage medium for executing the method. Background Art

[0002] This section is intended to provide background or context for the embodiments of the invention set forth in the claims. The descriptions herein may include concepts that may be explored, but not necessarily concepts that have been previously thought of or explored. Therefore, unless otherwise noted herein, the content described in this section is not prior art with respect to the specification and claims of this application, and is not admitted to be prior art by inclusion in this section.

[0003] In the related art, when distributing audio content to different channels such as large-screen TVs, mobile terminals, or in-vehicle terminals, if it is necessary to add content to the distributed audio content, all the source files of the audio to be distributed need to be edited one by one, that is, the source files need to be re-recorded, and content needs to be added during the recording process, and then edited and added to the background for upload. It can be seen that the existing audio content distribution process is very cumbersome and inefficient. Summary of the invention

[0004] The known distribution of audio content is not ideal and can be a very frustrating process.

[0005] Therefore, there is a great need for an improved audio resource distribution solution that can effectively simplify the audio resource distribution process and improve resource distribution efficiency.

[0006] In this context, embodiments of the present invention are intended to provide a method for distributing audio resources and related products.

[0007] In a first aspect of an embodiment of the present invention, a method for distributing audio resources is proposed, comprising: determining a distribution channel and additional resources of the audio resources in response to a demand for distribution of the audio resources; acquiring structural information about the audio resources; adding the additional resources to the audio resources according to the structural information of the audio resources to obtain resources to be distributed; and distributing the resources to be distributed through the distribution channel.

[0008] In one embodiment of the present invention, obtaining the structural information about the audio resource includes: obtaining structural segmentation granularity information of the audio resource; and performing structural processing on the audio resource according to the structural segmentation granularity information to obtain the structural information of the audio resource.

[0009] In another embodiment of the present invention, structuring the audio resource based on the structural segmentation granularity information includes: obtaining the original audio content of the audio resource; obtaining the text content corresponding to the original audio content, wherein the text content is time-related to the original audio content; analyzing the original audio content based on the structural segmentation granularity information to determine the boundary segmentation information of the original audio content; analyzing the text content based on the structural segmentation granularity information to determine the boundary segmentation time information of the original audio content; and determining the structural information of the audio resource based on the boundary segmentation information and the boundary segmentation time information of the original audio content.

[0010] In another embodiment of the present invention, structuring the audio resource based on the structural segmentation granularity information includes: obtaining the original audio content of the audio resource; obtaining the text content corresponding to the original audio content, wherein the text content is time-related to the original audio content; analyzing the text content based on the structural segmentation granularity information to determine the boundary segmentation information and boundary segmentation time information of the original audio content; and determining the structural information of the audio resource based on the boundary segmentation information and boundary segmentation time information of the original audio content.

[0011] In still another embodiment of the present invention, determining the structural information of the audio resource includes: fusing the boundary segmentation information and the boundary segmentation time information of the original audio content to obtain fused boundary information, wherein the boundary segmentation information and the boundary segmentation time information in the fused boundary information correspond one to one; and generating the structural information of the audio resource based on the fused boundary information.

[0012] In one embodiment of the present invention, generating the structural information of the audio resource based on the fused boundary information includes: performing clustering processing on the fused boundary information to determine the paragraph category of the paragraph to which the boundary segmentation information in the fused boundary information belongs; and generating the structural information of the audio resource based on the boundary segmentation information in the fused boundary information, and its corresponding boundary segmentation time information and / or corresponding paragraph category.

[0013] In another embodiment of the present invention, adding the additional resource to the audio resource according to the structural information of the audio resource includes: determining an additional position of the additional resource in the audio resource according to the structural information of the audio resource; and adding the additional resource at the additional position in the audio resource.

[0014] In another embodiment of the present invention, wherein the distribution requirement includes at least a channel identifier, determining the distribution channel and additional resources of the audio resource includes: determining the distribution channel of the audio resource according to the channel identifier; and acquiring additional resources related to the distribution channel.

[0015] In yet another embodiment of the present invention, the audio resource corresponds to one or more different distribution channels, and for each of the distribution channels, the method further comprises: displaying additional resources corresponding to each of the distribution channels and their additional positions in the audio resources.

[0016] In a second aspect of the embodiments of the present invention, an electronic device is provided, comprising: a processor; and a memory storing computer instructions for distributing audio resources, wherein when the computer instructions are executed by the processor, the electronic device executes the method described in the foregoing and following embodiments.

[0017] In a third aspect of the embodiments of the present invention, a computer-readable storage medium is provided, comprising program instructions for distributing audio resources. When the program instructions are executed by a processor, the method according to the above and the following embodiments is implemented.

[0018] According to the audio resource distribution method and related products of the embodiment of the present invention, the obtained structural information of the audio resource can be used to add related additional resources to the audio resource to determine the resource to be distributed, and the resource to be distributed is distributed through the determined distribution channel to achieve resource distribution. It can be seen that when distributing audio resources with the need to add additional resources, the solution of the present invention can add content to any position of the audio resource based on the structural information of the audio resource without re-recording the source file of the audio resource. In this way, the distribution process of the audio resource can be effectively simplified and the distribution efficiency can be improved. BRIEF DESCRIPTION OF THE DRAWINGS

[0019] The above and other objects, features and advantages of the exemplary embodiments of the present invention will become readily understood by reading the following detailed description with reference to the accompanying drawings. In the accompanying drawings, several embodiments of the present invention are shown in an exemplary and non-limiting manner, in which:

[0020] Figure 1 A block diagram schematically illustrates an exemplary computing system 100 suitable for implementing embodiments of the present invention;

[0021] Figure 2 The following is a schematic diagram showing a flow chart of a method for distributing audio resources according to an embodiment of the present invention;

[0022] Figure 3The following is a schematic diagram showing a flow chart of a method for distributing audio resources according to another embodiment of the present invention;

[0023] Figure 4 The following is a schematic diagram showing a flow chart of a method for distributing audio resources according to another embodiment of the present invention;

[0024] Figure 5 A schematic diagram schematically shows a channel distribution management interface according to an embodiment of the present invention; and

[0025] Figure 6 The structure diagram of the electronic device according to the embodiment of the present invention is schematically shown.

[0026] In the drawings, the same or corresponding reference numerals represent the same or corresponding parts. DETAILED DESCRIPTION

[0027] The principles and spirit of the present invention will be described below with reference to several exemplary embodiments. It should be understood that these embodiments are provided only to enable those skilled in the art to better understand and implement the present invention, and are not intended to limit the scope of the present invention in any way. On the contrary, these embodiments are provided to make the present disclosure more thorough and complete, and to fully convey the scope of the present disclosure to those skilled in the art.

[0028] The principles and spirit of the present invention will be described below with reference to several exemplary embodiments. It should be understood that these embodiments are provided only to enable those skilled in the art to better understand and implement the present invention, and are not intended to limit the scope of the present invention in any way. On the contrary, these embodiments are provided to make the present disclosure more thorough and complete, and to fully convey the scope of the present disclosure to those skilled in the art.

[0029] Figure 1 1 is a block diagram of an exemplary computing system 100 suitable for implementing embodiments of the present invention. Figure 1As shown, the computing system 100 may include: a central processing unit (CPU) 101, a random access memory (RAM) 102, a read-only memory (ROM) 103, a system bus 104, a hard disk controller 105, a keyboard controller 106, a serial interface controller 107, a parallel interface controller 108, a display controller 109, a hard disk 110, a keyboard 111, a serial external device 112, a parallel external device 113, and a display 114. Among these devices, the CPU 101, the RAM 102, the ROM 103, the hard disk controller 105, the keyboard controller 106, the serial controller 107, the parallel controller 108, and the display controller 109 are coupled to the system bus 104. The hard disk 110 is coupled to the hard disk controller 105, the keyboard 111 is coupled to the keyboard controller 106, the serial external device 112 is coupled to the serial interface controller 107, the parallel external device 113 is coupled to the parallel interface controller 108, and the display 114 is coupled to the display controller 109. It should be understood that Figure 1 The structural block diagram is only for the purpose of illustration, rather than limiting the scope of the present invention. In some cases, some equipment may be added or reduced according to specific circumstances.

[0030] It is known to those skilled in the art that the embodiments of the present invention may be implemented as a system, method or computer program product. Therefore, the present disclosure may be specifically implemented in the following forms, namely: complete hardware, complete software (including firmware, resident software, microcode, etc.), or a combination of hardware and software, generally referred to herein as a "circuit", "module", "unit" or "system". In addition, in some embodiments, the present invention may also be implemented in the form of a computer program product in one or more computer-readable media, which contains computer-readable program code.

[0031] Any combination of one or more computer-readable media may be used. A computer-readable medium may be a computer-readable signal medium or a computer-readable storage medium. A computer-readable storage medium may be, for example, but not limited to, an electrical, magnetic, optical, electromagnetic, infrared, or semiconductor system, device or device, or any combination of the above. More specific examples (non-exhaustive examples) of computer-readable storage media may include, for example: an electrical connection with one or more wires, a portable computer disk, a hard disk, a random access memory (RAM), a read-only memory (ROM), an erasable programmable read-only memory (EPROM or flash memory), an optical fiber, a portable compact disk read-only memory (CD-ROM), an optical storage device, a magnetic storage device, or any suitable combination of the above. In this document, a computer-readable storage medium may be any tangible medium containing or storing a program that may be used by or in combination with an instruction execution system, device or device.

[0032] Computer-readable signal media may include data signals propagated in baseband or as part of a carrier wave, which carry computer-readable program code. Such propagated data signals may take a variety of forms, including but not limited to electromagnetic signals, optical signals, or any suitable combination of the above. Computer-readable signal media may also be any computer-readable medium other than a computer-readable storage medium, which may send, propagate, or transmit a program for use by or in conjunction with an instruction execution system, apparatus, or device.

[0033] The program code embodied on the computer readable medium may be transmitted using any appropriate medium, including but not limited to wireless, wireline, optical fiber cable, RF, etc., or any suitable combination of the foregoing.

[0034] Computer program code for performing the operation of the present invention may be written in one or more programming languages ​​or a combination thereof, including object-oriented programming languages ​​such as Java, Smalltalk, C++, and conventional procedural programming languages ​​such as "C" or similar programming languages. The program code may be executed entirely on the user's computer, partially on the user's computer, as a separate software package, partially on the user's computer and partially on a remote computer, or entirely on a remote computer or server. In the case of a remote computer, the remote computer may be connected to the user's computer through any type of network, including a local area network (LAN) or a wide area network (WAN), or may be connected to an external computer (e.g., via the Internet using an Internet service provider).

[0035] The following will describe the implementation of the present invention with reference to the flowchart of the method of the embodiment of the present invention and the block diagram of the device (or system). It should be understood that each square frame of the flowchart and / or block diagram and the combination of square frames in the flowchart and / or block diagram can be implemented by computer program instructions. These computer program instructions can be provided to a processor of a general-purpose computer, a special-purpose computer or other programmable data processing device, thereby producing a machine, and these computer program instructions are executed by a computer or other programmable data processing device to produce a device that implements the function / operation specified in the square frame in the flowchart and / or block diagram.

[0036] These computer program instructions may also be stored in a computer-readable medium that enables a computer or other programmable data processing device to operate in a specific manner, so that the instructions stored in the computer-readable medium produce a product that includes an instruction device that implements the functions / operations specified in the blocks in the flowchart and / or block diagram.

[0037] Computer program instructions may also be loaded onto a computer, other programmable data processing apparatus, or other device so that a series of operational steps are performed on the computer, other programmable data processing apparatus, or other device to produce a computer-implemented process, thereby enabling the instructions executed on the computer or other programmable device to provide a process for implementing the functions / operations specified in the blocks in the flowchart and / or block diagram.

[0038] According to an embodiment of the present invention, a method for distributing audio resources and related products thereof are proposed. In addition, any number of elements in the drawings is for illustration rather than limitation, and any naming is only for distinction and does not have any limiting meaning.

[0039] The principle and spirit of the present invention are explained in detail below with reference to several representative embodiments of the present invention. SUMMARY OF THE INVENTION

[0041] The inventors have found that the distribution effect of existing audio resources is not ideal. Specifically, when distributing audio content, if you want to add new audio content to the audio content, you can only re-record the source file and record the new content in the process of recording the source file. For example, if you need to add new audio (such as a prompt voice) to the head of an audio about "Why gibbons have no tails", you need to re-record the audio containing the prompt voice. Especially when the source file is large or audio with different content needs to be distributed on different channels, the entire audio resource distribution process takes a long time and the operation is very cumbersome.

[0042] Based on this, the inventor has found through research that the audio resources can be structured and the structural information of the audio resources can be used to support adding additional resources at any position in the audio resources without re-editing the source file. This can greatly shorten the audio resource distribution cycle and effectively improve the distribution efficiency.

[0043] After introducing the basic principles of the present invention, various non-limiting embodiments of the present invention are described in detail below.

[0044] Exemplary Methods

[0045] Reference below Figure 2 The method for distributing audio resources according to an exemplary embodiment of the present invention is described below. It should be noted that the embodiments of the present invention can be applied to any applicable scenario.

[0046] Figure 2 The flowchart of a method 200 for distributing audio resources implemented by a computer according to an embodiment of the present invention is schematically shown.

[0047] like Figure 2As shown, at step S201, in response to the distribution demand for the audio resource, the distribution channel and additional resources of the audio resource are determined. The audio resource here can be understood as a resource including pure audio content or other multimedia resources with audio (such as audio and video mixed resources), etc. The specific type of the audio resource is not limited here and can be adjusted according to application requirements.

[0048] The aforementioned additional resources may include audio, text information, picture information or other types of information. The specific type and content of the additional resources are not limited here and can be adjusted according to application requirements and configuration information of the distribution channel.

[0049] In practical applications, the distribution channel and additional resources of the audio resource can be determined in a variety of ways. For example, in some embodiments, the channel information of the distribution channel of the audio resource and the corresponding additional resources can be pre-stored, and when it is determined that the distribution demand for the audio resource is obtained, the pre-stored channel information about the audio resource and the corresponding additional resources are called. In other embodiments, the channel information and the corresponding additional resources uploaded by the user in real time can also be obtained. In some other embodiments, the channel information of the distribution channel of the audio resource can also be pre-stored, and the pre-stored channel information is called according to the distribution demand of the audio resource, and the user is prompted to upload the corresponding additional information in real time according to the channel information. In some other embodiments, the distribution demand of the audio resource contains the distribution channel information and the additional resource information of the audio resource, and the corresponding distribution channel and the attachment resource can be obtained by parsing the distribution demand. It should be noted that the detailed description of the acquisition process of the distribution channel and the additional resources here is only an exemplary description, and can be adaptively adjusted according to the interaction design and application scenarios between the user.

[0050] Then, at step S202, the structural information of the audio resource can be obtained. The structural information here can be understood as the relevant information of the specific structure of the audio resource, which can be obtained by performing content structured segmentation on the audio resource. In some embodiments, the structural information of the audio resource can be pre-stored, and the structural information of the audio resource can be obtained by calling the pre-stored structural information. For example, in other embodiments, the structural information of the audio resource can also be obtained by performing content structured segmentation on the audio resource in real time. It should be noted that the detailed description of the structural information here is only an exemplary description, and the scheme of the present invention is not limited thereto. For example, for some newly released audio resources, when the audio resource is distributed for the first time, the real-time content structured segmentation can be performed for the audio resource, and then the structural information obtained by the segmentation can be stored, so that it can be directly called when there is a subsequent demand for the use of the structural information. As a result, there is no need to repeat the structural processing of the audio resource, and only the new audio resource needs to be parsed for the first time to obtain its structural information, which can be reused according to demand later.

[0051] Then, at step S203, the aforementioned additional resources can be added to the audio resource according to the structural information of the audio resource to obtain the resource to be distributed. As mentioned above, the structural information can be obtained by structurally segmenting the audio content of the audio resource. The structural information of the audio resource can be used to clarify the specific structural information of each part of the audio content of the audio resource, that is, the audio content in the audio resource may be structured into segments or even sentences, so that the additional resources can be added to any position in the audio resource (for example, a segment or a sentence) with the help of the structural information of the audio resource.

[0052] Finally, at step S204, the aforementioned resources to be distributed can be distributed through distribution channels. After obtaining the above-mentioned distribution resources, the resources to be distributed can be distributed to the corresponding distribution channels. The distribution channels here may include large-screen TV terminals, mobile terminals (including various audio and video playback platforms in the terminals) or vehicle-mounted terminals, etc. In some embodiments, the distribution channels can also be divided into internal channels and external channels, and the above-mentioned resources to be distributed are selectively distributed to different channels to meet different distribution needs. It should be noted that the description of the types and division methods of the distribution channels here is only an exemplary description.

[0053] Therefore, when distributing audio resources that require additional resources to be added, content can be added to any location of the audio resources based on the structural information of the audio resources without re-recording the source files of the audio resources. This can greatly shorten the audio resource distribution cycle, thereby effectively simplifying the audio resource distribution process and improving distribution efficiency.

[0054] Figure 3 The flowchart of the method 300 for distributing audio resources according to another embodiment of the present invention is schematically shown. It can be understood that the method 300 is for Figure 2 Therefore, the above combined with the method 200 Figure 2 The relevant detailed description also applies to the following.

[0055] like Figure 3 As shown, at step S301, the distribution channel of the audio resource can be determined according to the channel identifier in the distribution demand, and the additional resources of the distribution channel can be obtained. In some embodiments, the distribution demand initiated by the user for the audio resource will include a channel identifier (such as a channel name or other information that can distinguish the channel, etc.), and the channel identifier is obtained by parsing the distribution demand, and the distribution channel of the audio resource is determined according to the channel identifier. In practical applications, the additional resource can be obtained in a variety of ways. For example, an association relationship between a plurality of additional resources and corresponding channel identifiers is pre-stored, and in the pre-stored association relationship, the additional resource corresponding to the channel identifier included in the distribution demand is searched. For another example, the additional information corresponding to the distribution channel uploaded by the user in real time can also be obtained. For another example, the distribution demand initiated by the user for the audio resource can also include a channel identifier and a corresponding additional resource at the same time, and the distribution channel and additional resources of the audio resource can be determined by parsing the distribution demand.

[0056] Next, at step S302, the structure segmentation granularity information of the audio resource may be obtained. The structure segmentation granularity information may be used to indicate the granularity of the structured segmentation processing, such as segmentation by paragraph, line or sentence, and even word segmentation may be realized with the update of segmentation technology. In some embodiments, the structure segmentation granularity information may be set by default, or the user may adjust the specific segmentation granularity according to the specific application scenario.

[0057] Next, at step S303, the audio resource may be structured according to the structure segmentation granularity information to obtain the structure information of the audio resource. After obtaining the specific structure segmentation granularity information about the audio resource, the audio resource may be structured according to the granularity size represented by the structure segmentation granularity information, such as line by line or segment by segment.

[0058] In practical applications, there are many ways to implement the structured processing of audio resources. For example, in some embodiments, the text content corresponding to the original audio content can be obtained. Wherein, the text content is related to the original audio content timing, that is, the text content is compared with the content of the original audio content and the timestamps match. Then, the original audio content is analyzed according to the aforementioned structural segmentation granularity information to determine the boundary segmentation information of the original audio content, and the text content is analyzed according to the structural segmentation granularity information to determine the boundary segmentation time information of the original audio content.

[0059] Among them, the analysis of the original audio content can be performed by using traditional speech recognition analysis technology for recognition analysis, and the boundary segmentation information is determined in combination with the structural segmentation granularity information. For example, the original audio content can be segmented using traditional speech recognition technology, and then the above-mentioned segmentation processing result is determined to be segmented by paragraph or sentence by sentence in combination with the structural segmentation granularity information, and the boundary segmentation information is determined according to each segmentation position of the segmentation processing result. For text content analysis, the text can be segmented in combination with traditional lexical, syntactic and semantic analysis (of course, if the text content has clear structural information, the text segmentation processing can also be realized directly through the structural information of the text content), and the above-mentioned segmentation processing result is determined to be segmented by paragraph or sentence by sentence in combination with the structural segmentation granularity, so as to obtain the specific time corresponding to each segmentation position of the segmentation processing result to determine the boundary segmentation information. The text content matches the original audio content timestamp, and the boundary segmentation time information of the original audio content can be determined according to the boundary segmentation time information of the text content.

[0060] In other embodiments, the original audio content of the audio resource and the text content corresponding to the original audio content (i.e., the text content whose content and timestamp match the original audio content) can also be obtained. Then, the text content is analyzed according to the aforementioned structural segmentation granularity information to determine the boundary segmentation information and boundary segmentation time information of the original audio content. For example, the text content can be segmented using traditional text analysis techniques, and the above-mentioned segmentation processing results can be segmented segmented or segmented sentence by sentence in combination with the structural segmentation granularity to obtain the various segmentation positions of the segmentation processing results and the specific time corresponding to each segmentation position. Since the text content and the original text content match in terms of specific content and timestamp, the various segmentation positions of the segmentation processing results and the specific time corresponding to each segmentation position can be used to determine the various segmentation positions of the segmentation processing results and the specific time corresponding to each segmentation position.

[0061] Further, after obtaining the boundary segmentation information and boundary segmentation time information of the original audio content, the boundary segmentation information and the boundary segmentation time information can be fused to obtain the fused boundary information, and the structural information of the audio resource is determined according to the fused boundary information. Among them, the boundary segmentation information and the boundary segmentation time information in the fused boundary information correspond one to one. In some embodiments, the fused boundary information can be clustered to determine the paragraph category of the paragraph to which the boundary segmentation information in the fused boundary information belongs. Then, the structural information of the audio resource can be generated according to the boundary segmentation information in the fused boundary information, and its corresponding boundary segmentation time information and / or the corresponding paragraph category. For example, the structural information of the audio resource can specifically include each segmentation position and the specific time corresponding to each segmentation position, and can further include the specific paragraph corresponding to each segmentation position, etc. Thus, the structured processing of the audio resource is realized.

[0062] After completing the structural processing of the audio resource to obtain the structural information, at step S304, the additional position of the additional resource in the audio resource can be determined according to the structural information of the audio resource, and the additional resource can be added at the additional position in the audio resource to obtain the resource to be distributed. As shown in the foregoing, the structural information of the audio resource can include information such as each segmentation position and its corresponding specific paragraph. In some embodiments, the structural information of the audio resource can be displayed to the user for user selection, and the additional position of the additional resource in the audio resource is determined according to the segmentation position selected by the user. For another example, the distribution demand for the audio resource can also include the position information to be added of the additional position, and the segmentation position corresponding to the position information to be added is searched from the structural information of the audio resource, and the additional resource is added at the segmentation position found. For another example, the additional resource can also be analyzed and the specific additional position in the audio resource can be determined in combination with the preset addition rules. Specifically, taking the additional resource as copyright information as an example, the additional resource is determined to be copyright information by analyzing the additional resource. At this time, according to the predefined adding rule of "copyright information is added to the header of the audio resource", the header position of the audio resource is searched from the structural information of the audio resource, and the additional information is added to the found position.

[0063] Finally, at step S305, the aforementioned resources to be distributed may be distributed through a distribution channel, thereby achieving efficient distribution of audio resources.

[0064] Figure 4 The flowchart of the method 400 for distributing audio resources according to another embodiment of the present invention is schematically shown. It can be understood that the method 400 can be understood as a specific technical implementation of the method 200 or the method 300. Figure 2 and Figure 3The relevant detailed description in also applies to the following.

[0065] like Figure 4 As shown, at step S401, the audio content and text content of the content to be analyzed are obtained. Among them, the audio content and the text content match. The content to be analyzed here is the audio resource mentioned above, which can be pure audio or other multimedia resources such as video with voice. Whether it is pure audio or other multimedia resources with mixed voice, the audio content can be extracted from the audio resource. The text content corresponding to the audio content can be obtained by performing voice recognition processing on the audio content based on traditional voice recognition technology. Of course, it can also be text content uploaded by the user or pre-stored to correspond to the audio content. The description of the process of obtaining the text content here is only an exemplary description.

[0066] Next, at step S402, the audio content may be analyzed to determine the boundary segmentation information of the audio content. The boundary segmentation information here may include each segmentation position information. In some embodiments, the audio content may be segmented using conventional speech recognition technology, and the boundary segmentation information may be further determined based on the segmentation result and the specific structure division granularity.

[0067] At step S403, the above-mentioned text content can be analyzed to determine the boundary segmentation time information of the audio content. The boundary segmentation information here can be understood as the time corresponding to each segmentation position information. In certain embodiments, the text content can be segmented, and the boundary segmentation time information of the text content can be further determined according to the segmentation result of the text content and the specific structure division granularity. The text content corresponds to the audio content, and the boundary segmentation time information of the audio content can be determined according to the boundary segmentation time information of the text content.

[0068] At step S404, the boundary segmentation information and the boundary segmentation time information of the audio content may be fused to obtain fused boundary information, in which the segmentation position and time are accurately matched one by one.

[0069] At step S405, the fused boundary information may be clustered to determine the categories of each paragraph in the content to be analyzed, and further determine the structure of the content to be analyzed. Thus, the structural analysis of the content to be analyzed is achieved.

[0070] At step S406, the resource obtained through the structured processing (ie, the structured audio resource) may be stored in the background, and additional information may be selectively added and distributed according to distribution requirements.

[0071] like Figure 5As shown, unified management of structured audio can be achieved through background management platforms such as "Channel Content Pool Management". Specifically, the distribution channel can be selected through the "Channel Name" option, and whether to add additional information (such as copyright information, subsequent plot prompt information, marketing information, or information prompting users to join the group, etc.) can be selected through the "Whether to add additional information" option. The specific location for adding the additional information is selected through the "Added Location" option. For example, the structural information of the structured audio can be displayed to the user through the "Added Location", and the specific location for adding the additional information is determined according to the user's choice. The source file corresponding to the structured audio is uploaded through the "Added Source File" option. It should be noted that the detailed description of the background management platform here is only an exemplary description. In actual applications, a management platform that supports related functions can be built according to actual interactive design and application scenarios.

[0072] Further, it is also possible to customize the addition of additional information in batches for different additional contents and distribution channels. Specifically, when the audio resource corresponds to one or more different distribution channels, for each distribution channel, the additional resources corresponding to each distribution channel and its additional position in the audio resource can be displayed for user confirmation and / or adjustment, thereby realizing resource batch distribution management. In some embodiments, the audio resource corresponds to different distribution channels, such as internal channels and external channels, etc. When it is necessary to distribute the audio resource on the internal channel, the source file without adding additional information can be directly distributed. Of course, it is also supported to publish the audio resource with added additional information on the internal channel. When it is necessary to distribute the audio resource on the external channel, the audio resource with the same additional information or the audio resource with different additional information can be distributed to each external channel. Thus, by segmenting the audio resource in a structured manner, it is supported to add any required additional information at any position of the audio resource, and it can be selected whether to distribute these audio resources with attached information according to different channels.

[0073] Exemplary Devices

[0074] After introducing the method of the exemplary embodiment of the present invention, next, refer to Figure 6 The related products of the method for distributing audio resources according to the exemplary embodiment of the present invention are described.

[0075] Figure 6 Schematically shows a schematic block diagram of an electronic device 600 according to an embodiment of the present invention. Figure 6 As shown, the electronic device 600 may include a processor 601 and a memory 602. The memory 602 stores a computer instruction for distributing audio resources. When the computer instruction is executed by the processor 601, the electronic device 600 executes the above-mentioned method in combination with the above-mentioned method. Figures 2 to 4 The method described. For example, in some embodiments, the electronic device 600 can obtain the distribution channel and additional resources of the audio resource, implement the structured analysis of the audio resource and obtain the structure information, etc. Based on this, when distributing an audio resource with the requirement of adding additional resources, the electronic device 600 can add content to any position of the audio resource according to the structure information of the audio resource, without having to re-record the source file of the audio resource to add the content. In this way, the distribution process of the audio resource can be effectively simplified and the distribution efficiency can be improved.

[0076] It should be noted that although several devices or sub-devices of the device are mentioned in the above detailed description, this division is not mandatory. In fact, according to an embodiment of the present invention, the features and functions of two or more devices described above can be embodied in one device. Conversely, the features and functions of one device described above can be further divided into multiple devices to be embodied.

[0077] The use of the verbs "comprise", "include" and their conjugations mentioned in the application documents does not exclude the presence of elements or steps other than those recorded in the application documents. The article "a" or "an" before an element does not exclude the presence of a plurality of such elements.

[0078] Although the spirit and principle of the present invention have been described with reference to several specific embodiments, it should be understood that the present invention is not limited to the disclosed specific embodiments, and the division of various aspects does not mean that the features in these aspects cannot be combined to benefit, and this division is only for the convenience of expression. The present invention is intended to cover various modifications and equivalent arrangements included in the spirit and scope of the attached claims. The scope of the attached claims conforms to the broadest interpretation, thereby including all such modifications and equivalent structures and functions.

Claims

1. A method for distributing audio resources, characterized in that: include: In response to a demand for distribution of the audio resource, determining a distribution channel and additional resources for the audio resource; Obtaining structural information about the audio resource; Determining, according to the structural information of the audio resource, an additional position of the additional resource in the audio resource; adding the additional resource at the additional position in the audio resource to obtain the resource to be distributed; and Distributing the resources to be distributed through the distribution channel; The distribution requirement includes the location information of the additional location to be added; the structural information includes each segmentation location; and determining the additional location of the additional resource in the audio resource according to the structural information of the audio resource includes: searching the structure information for a segmentation position corresponding to the position information to be added; or, Analyze the additional resource and determine the specific additional position of the additional resource in the audio resource in combination with a preset adding rule; Wherein, adding the additional resource at the additional position in the audio resource comprises: adding the additional resource at the found segmentation position; or, The additional resource is added at the specific additional location.

2. The distribution method according to claim 1, characterized in that: Acquiring structural information about the audio resource includes: Acquire structural segmentation granularity information of the audio resource; and The audio resource is structured according to the structural segmentation granularity information to obtain structural information of the audio resource.

3. The distribution method according to claim 2, characterized in that: The structural processing of the audio resource according to the structural segmentation granularity information includes: Obtaining the original audio content of the audio resource; Acquire text content corresponding to the original audio content, wherein the text content is time-sequentially related to the original audio content; Analyzing the original audio content according to the structural segmentation granularity information to determine boundary segmentation information of the original audio content; Analyzing the text content according to the structural segmentation granularity information to determine boundary segmentation time information of the original audio content; and The structural information of the audio resource is determined according to the boundary segmentation information and the boundary segmentation time information of the original audio content.

4. The distribution method according to claim 2, characterized in that: The structural processing of the audio resource according to the structural segmentation granularity information includes: Obtaining the original audio content of the audio resource; Acquire text content corresponding to the original audio content, wherein the text content is time-sequentially related to the original audio content; Analyzing the text content according to the structural segmentation granularity information to determine boundary segmentation information and boundary segmentation time information of the original audio content; and The structural information of the audio resource is determined according to the boundary segmentation information and the boundary segmentation time information of the original audio content.

5. The distribution method according to claim 3 or 4, characterized in that: Determining the structural information of the audio resource includes: Fusing the boundary segmentation information and the boundary segmentation time information of the original audio content to obtain fused boundary information, wherein the boundary segmentation information and the boundary segmentation time information in the fused boundary information correspond to each other one by one; and The structure information of the audio resource is generated according to the fusion boundary information.

6. The distribution method according to claim 5, characterized in that: Generating the structure information of the audio resource according to the fusion boundary information includes: performing clustering processing on the fused boundary information to determine the paragraph category of the paragraph to which the boundary segmentation information in the fused boundary information belongs; and The structural information of the audio resource is generated according to the boundary segmentation information in the fused boundary information, and its corresponding boundary segmentation time information and / or corresponding paragraph category.

7. The distribution method according to claim 1, characterized in that: The distribution requirement at least includes a channel identifier, and determining the distribution channel and additional resources of the audio resource includes: Determining a distribution channel of the audio resource according to the channel identifier; and Obtain additional resources related to the distribution channel.

8. The distribution method according to claim 7, characterized in that: The audio resource corresponds to one or more different distribution channels, and for each of the distribution channels, the method further includes: The additional resources corresponding to each of the distribution channels and their additional positions in the audio resources are displayed.

9. An electronic device, characterized in that: include: processor; as well as A memory storing computer instructions for distributing audio resources, wherein when the computer instructions are executed by the processor, the electronic device executes the method according to any one of claims 1 to 8.

10. A computer-readable storage medium, characterized in that: The program instructions comprising the distribution of audio resources, when the program instructions are executed by a processor, enable the method according to any one of claims 1-8 to be implemented.

Citation Information

Patent Citations

  • Music structure determination method and device, equipment and medium

    CN112037764A

  • Method and system for inserting audio or video

    CN114925223A