Audio playing method and device, electronic equipment and storage medium
By acquiring multi-dimensional target data, intelligently determining the audio playback device, the problem of inflexible device selection in the prior art is solved, and the intelligence level of the audio playback system is improved.
Patent Information
- Application Number
- CN202510609618.6
- Authority / Receiving Office
- CN · China
- Patent Type
- Applications(China)
- Current Assignee / Owner
- Filing Date
- 2025-05-13
- Publication Date
- 2025-06-10
- Estimated Expiration
- 2045-05-13
AI Technical Summary
The existing audio playback system lacks flexibility when selecting playback devices, and users need to perform complicated device selection operations, which is less intelligent.
By obtaining multi-dimensional target data, including location information, device configuration information, historical control information, audio attribute information and current space-time information, the target audio playback device is intelligently determined.
It realizes intelligent selection of audio playback equipment, improves the intelligence of the audio playback system, and reduces the operation steps of users to manually select devices.
Smart Images

Figure CN120122912A_ABST
Abstract
Description
Technical Field
[0001] The present disclosure relates to the technical field of audio playback, and in particular, to a method, apparatus, electronic device, and storage medium for playing audio. Background Art
[0002] In a playback system configured with multiple playback devices, when determining which one or more playback devices to use for each audio playback, usually one or more playback devices pre-specified by the user are quickly set as the default devices for each playback through default configuration, or the playback devices are specified through the user's operation. This method is not flexible. If the playback device that the user wants to use for the current playback is not the default device, complicated playback device selection operations are required, and the intelligence level is low. Summary of the Invention
[0003] In view of this, an object of the present disclosure is to provide a method, apparatus, electronic device, and storage medium for playing audio, so as to improve the intelligence level of determining an audio playback device.
[0004] In a first aspect, an embodiment of the present disclosure provides a method for playing audio, which is applied to an audio playback system. The audio playback system includes at least two audio playback devices. The method includes: in response to a playback instruction for a target audio, obtaining at least one piece of target data; where the at least one piece of target data includes any item of the position information of a detection target, the configuration information of each audio playback device in the audio playback system, the historical control information of the audio playback system, the attribute information of the target audio, and the current spatio-temporal information where the audio playback system is located; determining at least one target audio playback device from the audio playback system according to the at least one piece of target data; and playing the target audio through the target audio playback device.
[0005] In a second aspect, an embodiment of the present disclosure provides an apparatus for playing audio, which is applied to an audio playback system. The audio playback system includes at least two audio playback devices. The apparatus includes: an obtaining module, configured to obtain at least one piece of target data in response to a playback instruction for a target audio; where the at least one piece of target data includes any item of the position information of a detection target, the configuration information of each audio playback device in the audio playback system, the historical control information of the audio playback system, the attribute information of the target audio, and the current spatio-temporal information where the audio playback system is located; a determining module, configured to determine at least one target audio playback device from the audio playback system according to the at least one piece of target data; and a playing module, configured to play the target audio through the target audio playback device.
[0006] In a third aspect, embodiments of the present disclosure provide an electronic device, including a processor and a memory. The memory stores machine-executable instructions that can be executed by the processor, and the processor executes the machine-executable instructions to implement the above-mentioned audio playing method.
[0007] In a fourth aspect, embodiments of the present disclosure provide a computer-readable storage medium. The computer-readable storage medium stores computer-executable instructions, and when the computer-executable instructions are called and executed by a processor, the computer-executable instructions cause the processor to implement the above-mentioned audio playing method.
[0008] Embodiments of the present disclosure bring the following beneficial effects: The above-mentioned audio playing method, device, electronic device and storage medium predict an audio playing device through target data in different dimensions, so that the audio playing can intelligently select a playing device, rather than simply presetting a playing device through default settings. The intelligence level of the audio playing system is improved.
[0009] Other features and advantages of the present disclosure will be described in the following specification, and part of them will become obvious from the specification or be understood by implementing the present disclosure. The objectives and other advantages of the present disclosure are achieved and obtained by the structures specifically pointed out in the specification, claims and drawings.
[0010] To make the above objectives, features and advantages of the present disclosure more obvious and understandable, the following specific preferred embodiments are given and described in detail in conjunction with the accompanying drawings. Description of the Drawings
[0011] In order to more clearly illustrate the specific embodiments of the present disclosure or the technical solutions in the prior art, the following will briefly introduce the drawings required for the description of the specific embodiments or the prior art. Obviously, the following drawings are some embodiments of the present disclosure. For those skilled in the art, other drawings can be obtained based on these drawings without creative efforts.
[0012] Figure 1 It is a flowchart of an embodiment of the audio playing method in the embodiments of the present disclosure; Figure 2 It is a schematic diagram of an audio playing device provided by the embodiments of the present disclosure; Figure 3 It is a schematic diagram of an electronic device provided by the embodiments of the present disclosure. Detailed Embodiments
[0013] To make the objectives, technical solutions, and advantages of the embodiments of the present disclosure clearer, the technical solutions of the present disclosure will be clearly and completely described below with reference to the accompanying drawings. Apparently, the described embodiments are some but not all of the embodiments of the present disclosure. Based on the embodiments of the present disclosure, all other embodiments obtained by those of ordinary skill in the art without creative efforts shall fall within the scope of protection of the present disclosure.
[0014] The terms "first", "second", "third", "fourth", etc. (if any) in the specification, claims, and the above-mentioned drawings of the present disclosure are used to distinguish similar objects and do not necessarily describe a specific order or sequence. It should be understood that the data used in this way can be interchanged under appropriate circumstances so that the embodiments described here can be implemented in an order different from that shown or described here. In addition, the terms "comprising" or "having" and any variations thereof are intended to cover non-exclusive inclusion. For example, a process, method, system, product, or device that includes a series of steps or units does not necessarily limit to those clearly listed steps or units, but may include other steps or units not clearly listed or inherent to these processes, methods, products, or devices.
[0015] For ease of understanding, the specific process of the embodiments of the present disclosure will be described below. The embodiments of the present disclosure are applied to an audio playback system, and the audio playback system includes at least two audio playback devices. Please refer to Figure 1 , and an embodiment of the audio playback method in the embodiments of the present disclosure includes: Step S10, in response to a playback instruction for a target audio, obtain at least one piece of target data; wherein, the at least one piece of target data includes any item of the location information of the detection target, the configuration information of each audio playback device in the audio playback system, the historical control information of the audio playback system, the attribute information of the target audio, and the current spatio-temporal information where the audio playback system is located; The user can trigger a playback instruction for the target audio through a terminal device. For example, the user triggers a playback instruction for the target audio through a television, a control panel, a mobile device (such as a mobile phone), etc. in the audio playback system, so that the target audio is played through the audio playback devices in the audio playback system. The playback instruction for the target audio can also be automatically triggered, such as being triggered regularly, etc., and specific details are not limited here.
[0016] The target data is data used to determine the target audio playback device. In this embodiment, one item of target data can be obtained, or more than one item of target data can be obtained. The target data can be any one or more of the position information of the detection target, the configuration information of each audio playback device in the audio playback system, the historical control information of the audio playback system, the attribute information of the target audio, and the current spatio-temporal information where the audio playback system is located. It can also be other data that can be used to determine the target audio playback device as listed, and specific details are not limited here.
[0017] The detection target can be any object whose position can be detected. By way of example and not limitation, the detection target can be a device, a living organism, or a designated marker. For example, the detection target can be a mobile phone, a bracelet, a watch, etc. that can be used to represent the position of the wearer, or any human body, a human body with facial features that meet a predetermined standard, a human body with voiceprint features that meet a predetermined standard, a human body with fingerprint features that meet a predetermined standard, etc. that can be used to represent the position of the human body. It can also be a marker with specific specific markings such as a predetermined icon, a predetermined shape, or a predetermined weight. Specific details are not limited here.
[0018] For a detection target of a device type, the position information of the detection target can be obtained through methods such as satellite positioning, WiFi signal detection, Bluetooth Low Energy (BLE) signal strength detection, etc. Among them, WiFi signal detection and BLE signal strength detection have higher accuracy in indoor positioning and can improve the accuracy of determination by the audio playback device.
[0019] The configuration information of the audio playback device can include the hardware configuration information, software configuration information, and related attribute information of the audio playback device. For example, the hardware configuration information can include the model of the digital-to-analog converter (DAC), the output power and signal-to-noise ratio of the amplifier, the parameters of the interface, the frequency response range, the impedance, the sensitivity, etc. The software configuration information can include the parameters of the operating system, the version of the driver, the sampling rate, the bit depth, the audio enhancement, the network streaming media protocol, the equalizer, etc. The related attribute information can include the identifier, the name, the spatial information where it is located, whether it is connected to an audio-visual device, etc. Specific details are not limited here.
[0020] The historical control information of the audio playback system includes the control information for controlling the audio playback system within a preset duration in the past. For example, the historical control information can include that at a certain moment in the past, a certain user triggered the playback of a certain audio on a certain audio playback device through a certain control device / method. Specific details are not limited here.
[0021] The attribute information of the target audio can include attributes related to sound quality, content type, style, or other attributes that can be used for selecting an audio playback device. The attribute information of the target audio can also be divided into technical attributes and perceptual attributes. For example, technical attributes can include sample rate (SampleRate), bit depth (Bit Depth), number of channels (Channels), bitrate (Bitrate), audio format (File Format), and frequency response (Frequency Response), etc. Perceptual attributes can include pitch, loudness, timbre, spatialization, clarity, etc. The specific details are not limited here.
[0022] The current spatio-temporal information refers to the spatio-temporal information of the environment where the audio playback system is located, which can include current time information, current weather information, current network environment information, current ambient light information, current ambient sound information, etc., which are information that can be used to represent the current environmental state. The specific details are not limited here.
[0023] Step S20: Determine at least one target audio playback device from the audio playback system according to at least one item of target data; In this embodiment, according to each item of the obtained target data, at least one audio playback device matching the target data can be determined from the audio playback system. In one embodiment, the matching degree between each item of target data and each audio playback device can be calculated respectively to obtain the first matching degree of each audio playback device corresponding to each item of target data, and then all the first matching degrees corresponding to each audio playback device are fused to obtain the target matching degree of each audio playback device. The audio playback device with a target matching degree higher than the preset matching degree threshold can be determined as the target audio playback device.
[0024] For example, if the audio playback system includes audio playback device A and audio playback device B, and the obtained target data includes target data A and target data B, then the first matching degrees of audio playback device A with target data A and target data B respectively, and the first matching degrees of audio playback device B with target data A and target data B respectively can be calculated first. Then, the first matching degrees of audio playback device A with target data A and target data B are fused to obtain the target matching degree of audio playback device A, and the first matching degrees of audio playback device B with target data A and target data B are fused to obtain the target matching degree of audio playback device B.
[0025] Next, determine whether the target matching degree of the audio playback device A and the target matching degree of the audio playback device B are higher than the preset matching degree threshold respectively. If the target matching degree of the audio playback device A is higher than the preset matching degree threshold, then determine the audio playback device A as the target audio playback device. The same applies to the audio playback device B, and specific details are not limited here.
[0026] In one implementation, according to the above target data, through a preset device feature generation algorithm, generate audio playback device feature data that matches the above target data, and then determine at least one target audio playback device that matches the above audio playback device feature data from the audio playback system according to the audio playback device feature data.
[0027] Step S30: Play the target audio through the target audio playback device.
[0028] After determining the target audio playback device, the target audio playback device can be controlled to play the target audio, so that the audio playback can intelligently select the playback device, rather than just intelligently preset the playback device through the default settings, and the intelligence level of the audio playback system is improved.
[0029] In one implementation, the playback parameters of the audio can also be determined according to the target data. When playing the target audio through the target audio playback device, play according to the determined playback parameters. Among them, the playback parameters can include parameters such as volume, equalizer, buffer size, and mode, and specific details are not limited here.
[0030] The audio playback method provided by the above implementation predicts the audio playback device through target data in different dimensions, so that the audio playback can intelligently select the playback device, rather than just intelligently preset the playback device through the default settings, and the intelligence level of the audio playback system is improved.
[0031] Next, a specific description of the method for determining the audio playback device for each item of target data will be given.
[0032] In one implementation, at least one item of target data is: the position information of the detection target, and the detection target is a target device or a target organism; when determining at least one target audio playback device from the audio playback system according to at least one item of target data, it includes: calculating the proximity between the detection target and the preset audible range of each audio playback device in the above audio playback system according to the position information; wherein, the preset audible range of each audio playback device is used to indicate the target communication space where the corresponding audio playback device is located; determine at least one target audio playback device whose proximity is greater than or equal to the preset proximity threshold from the audio playback system.
[0033] In the case where the target data is the location information of the detection target, when determining the target audio playback device, first, according to the location information of the detection target, calculate the proximity between the detection target and the audible regions of each audio playback device, and then determine the audio playback devices with a proximity greater than or equal to the preset proximity threshold as the target audio playback devices.
[0034] The preset audible region of an audio playback device refers to the region where the loudness of the sound is greater than or equal to the preset loudness threshold at a specific volume (sound pressure level). That is to say, when playing audio at a certain volume, if the loudness of the sound measured at a certain position is less than the above-mentioned preset loudness threshold, then that position does not belong to the preset audible region. The preset audible region is related to the position where the audio playback device is placed and the connectivity of the space it is in.
[0035] For example, assume an audio playback system in a home environment, and there is an audio playback device A placed in room A. Assume that when this audio playback device A plays audio at a certain volume in this room, the loudness heard at each position in the room is greater than or equal to the preset loudness threshold, while at any position outside this room, the detected loudness is less than the preset loudness threshold. Then, this room A is the preset audible region of this audio playback device A, and specific details are not limited here.
[0036] When calculating the proximity between the detection target and the preset audible regions of each audio playback device in the audio playback system, the proximity corresponding to the detection target within the preset audible region of the audio playback device can be determined as 100%. For the detection target outside the preset audible region of the audio playback device, the proximity can be calculated according to the distance between the detection target and the nearest boundary position of the preset audible region. For example, for every 10 cm increase in the distance, the proximity decreases by 1%, and specific details are not limited here.
[0037] The detection target can be a target device or a target organism. For example, it can be a target mobile phone, a target bracelet, a target watch, a human body, a target human body, etc. The target audio playback device determined by the location of the detection target can enable the playback of audio to sense the location of the target, achieving the effect of seamless control of playback and having a high degree of intelligence.
[0038] In one implementation, the proximity between the detection target and the preset audible region of the i-th audio playback device in the audio playback system can also be calculated through a preset proximity calculation formula :
[0039] where d refers to the distance between the detection target and the preset audible region of the i-th audio playback device in the audio playback system, refers to the maximum effective distance, and d exceeds the proximity degree is 0. When ≥ the preset proximity threshold, the i-th audio playback device can be determined as the target audio playback device.
[0040] For example, assume there are three smart speakers in a family, with speaker numbers 1, 2, and 3 in sequence, located in the living room (i = 1), bedroom (i = 2), and kitchen (i = 3) respectively. When the user wears a smart watch and enters the living room, the system calculates the distances between the user and the preset audible ranges of the three speakers through WiFi signal strength detection as follows: = 2 meters, = 8 meters, = 5 meters. Set = 10 meters, and the preset proximity threshold is 60%.
[0041] Then, the proximity degrees corresponding to the three smart speakers are: = 100% × (1 - 2 / 10) = 80% = 100% × (1 - 8 / 10) = 20% = 100% × (1 - 5 / 10) = 50% Since > 60%, determine the speaker located in the living room (i.e., the first speaker) as the target audio playback device. When the user issues a voice command of "Play today's news", the system automatically plays through the first speaker without the user manually selecting the device.
[0042] This algorithm realizes intelligent device selection based on the user's location, reduces the operation steps of the user manually selecting the device, and improves the usability and user experience of the system. At the same time, the algorithm uses relative distance percentage calculation, adapts to application scenarios of different space sizes, and has better generality.
[0043] In one implementation, at least one piece of target data is: the configuration information of each audio playback device in the audio playback system; when determining at least one target audio playback device from the audio playback system according to at least one piece of target data, it includes: obtaining the audio-visual type of the target audio; where the audio-visual type is used to indicate whether the target audio is a video and audio or a non-video and audio; calculating the device adaptation degrees of each audio playback device in the audio playback system according to the audio-visual type and the configuration information of each audio playback device; and determining the audio playback devices with device adaptation degrees greater than the preset device adaptation degree threshold as the target audio playback devices to obtain at least one target audio playback device.
[0044] The target audio can be the audio of a film or television, or it can be audio that is not from a film or television. For film and television audio and non-film and television audio, the device adaptability of the target audio can be calculated based on the configuration information of different audio playback devices to obtain a target audio playback device of a video and audio type that better suits the target audio.
[0045] When calculating the device adaptability of each audio playback device, for the target audio with the video and audio type of film and television audio, the number of configuration parameters that meet the video and audio playback requirements in the configuration information of each audio playback device can be calculated, and the device adaptability of each audio playback device can be determined based on the number of configuration parameters; for the target audio with the video and audio type of non-film and television audio, the number of configuration parameters that meet the target audio sound quality in the configuration information of each audio playback device can be calculated, and then the device adaptability of each audio playback device can be determined based on the number of configuration parameters.
[0046] When determining the device adaptability of each audio playback device based on the number of configuration parameters, the ratio between the number of configuration parameters and a preset number threshold or the maximum number of parameters can be calculated, and the ratio can be determined as the device adaptability of the corresponding audio playback device. Here, the maximum number of parameters refers to the maximum value among the number of configuration parameters, and specific details are not limited here.
[0047] In one implementation, the configuration adaptability between different video and audio types and different configuration parameters can be set through a preset corresponding relationship. When calculating the device adaptability of each audio playback device in the audio playback system based on the video and audio type and the configuration information of each audio playback device, the configuration adaptability between each configuration parameter in the configuration information of each audio playback device and the video and audio type of the target audio can be obtained from the above preset corresponding relationship first, and then the device adaptability of the m-th audio playback device in the audio playback system can be calculated through a preset device adaptability calculation formula :
[0048] Among them, represents the weight of the i-th configuration parameter of the m-th audio playback device, represents the configuration adaptability between the i-th configuration parameter of the m-th audio playback device and the video and audio type of the target audio.
[0049] For example, assume that the user requests to play the audio of a movie, and the video and audio type of this target audio is film and television audio. Assume that there are three smart speakers in the family, and the speaker numbers are 1, 2, and 3 in sequence, located in the living room (m = 1), bedroom (m = 2), and kitchen (m = 3) respectively. The configuration parameters of each smart speaker include the number of channels (i = 1), audio decoding ability (i = 2), power output (i = 3), and network connection performance (i = 4) in sequence. The weights of each configuration parameter correspond in sequence as: = 0.4, = 0.3, = 0.2, = 0.1.
[0050] Through the above preset correspondence, the configuration adaptation degrees between the configuration parameters of each smart speaker and the film and television audio are obtained as follows: Smart speaker No. 1 in the living room: 5.1-channel ( = 95%), Dolby decoding ( = 90%), 100W output ( = 85%), wired connection ( = 95%); Smart speaker No. 2 in the bedroom: 2.0-channel ( = 60%), AAC decoding ( = 70%), 20W output ( = 60%), Bluetooth connection ( = 75%); Smart speaker No. 3 in the kitchen: mono-channel ( = 40%), basic decoding ( = 50%), 10W output ( = 40%), WiFi connection ( = 90%).
[0051] Furthermore, the device adaptation degrees corresponding to the 3 smart speakers calculated are as follows: = (0.4 × 95% + 0.3 × 90% + 0.2 × 85% + 0.1 × 95%) / 1 = 91.5%; = (0.4 × 60% + 0.3 × 70% + 0.2 × 60% + 0.1 × 75%) / 1 = 64.5%; = (0.4 × 40% + 0.3 × 50% + 0.2 × 40% + 0.1 × 90%) / 1 = 48%; Assume that the preset device adaptation degree threshold is 70%. Then, since = 91.5% > 70%, therefore, the smart speaker in the living room is determined as the target audio playback device. Through multi-parameter weighted calculation, this algorithm realizes the precise matching of audio types and device configurations, ensuring that users obtain the best sound quality experience. The parameter weights of the algorithm can be dynamically adjusted according to different audio types, realizing the self-adaptability of the system, and at the same time improving the professionalism of audio playback and user satisfaction.
[0052] In one implementation, at least one piece of target data is: historical control information of the audio playback system; when determining at least one target audio playback device from the audio playback system according to at least one piece of target data, it includes: obtaining the current time information and the style label information of the target audio; according to the current time information and the style label information, matching at least one target audio playback device with a historical preference degree higher than the preset preference degree threshold from the historical control information.
[0053] The style label information of the target audio is used to indicate the style of the target audio. For example, classical style, pop style, rock style, jazz style, movie and TV style, etc. According to the current time information and the style label information, the target audio playback device that conforms to the historical preference can be matched from the historical control information, so that the user's habit of playing a specific style of audio at a specific time can be replicated, and the user experience in audio playback can be improved.
[0054] When matching at least one target audio playback device with a historical preference degree higher than the preset preference degree threshold from the historical control information according to the current time information and the style label information, it is possible to judge whether there is matching control information in the historical control information according to the current time information and the style label information. If so, the historical preference degree of each audio playback device can be calculated according to the matching control information, and then the audio playback device with a historical preference degree greater than the preset preference degree threshold is determined as the target audio playback device.
[0055] When judging whether there is matching control information in the historical control information, it is possible to judge whether there is control information with the same preset time period as the current time and the same style label in the historical control information. If so, the audio playback device that plays the corresponding control information is determined as the preferred playback device. When calculating the historical preference degree, the proportion of the number of each preferred playback device in the total number of all preferred playback devices is counted to obtain the historical preference degree corresponding to each preferred playback device. For non-preferred playback devices, their historical preference degree can be 0, and specific details are not limited here.
[0056] In one implementation, at least one piece of target data is: attribute information of the target audio; when determining at least one target audio playback device from the audio playback system according to at least one piece of target data, it includes: determining the content type of the target audio according to the attribute information; according to the content type, determining at least one target audio playback device that matches the preset adapted content type in the audio playback system.
[0057] The content type of the target audio refers to the type of text content in the target audio. According to whether there is text in the target audio, it can be divided into the text type and the non-text type. The non-text type is usually pure music or white noise, etc. The text type can be further divided into the dialogue type or the non-dialogue type. The dialogue type is usually the audio of a movie or TV drama, and the non-dialogue type is usually a song. For the non-dialogue type, the content type can also be divided according to the emotional tendency of the text content, such as the sad type, the cheerful type, the angry type, the plain type, etc. The specific division is not limited here.
[0058] After determining the content type of the target audio, the content type of the target audio can be matched with the preset adaptation content types corresponding to each audio playback device, and the audio playback devices that contain the same preset adaptation content type are determined as the target audio playback devices, so as to obtain at least one target audio playback device. Through content type matching, the playback device can have the type memory ability, avoid playing audio with inconsistent types on inappropriate playback devices, and improve the user experience.
[0059] In one implementation, at least one piece of target data is: the current spatio-temporal information where the audio playback system is located, and the current spatio-temporal information includes the current weather information and / or the current time information; when determining at least one target audio playback device from the audio playback system according to at least one piece of target data, it includes: predicting the behavior habit characteristics and emotional characteristics according to the current spatio-temporal information to obtain a characteristic prediction result; matching at least one target audio playback device whose characteristic matching degree with the characteristic prediction result is greater than the preset characteristic matching degree threshold according to the preset characteristic information of each audio playback device in the audio playback system.
[0060] When determining at least one target audio playback device from the audio playback system according to the current spatio-temporal information where the audio playback system is located, the behavior habits and emotions of the user can be predicted according to the current weather information and / or the current time information in the current spatio-temporal information to obtain the current behavior habit characteristics and emotional characteristics of the user, that is, the characteristic prediction result.
[0061] Then, the characteristic prediction result is matched with the preset characteristic information corresponding to each audio playback device, the characteristic matching degree is calculated to obtain the target characteristic matching degree corresponding to each audio playback device, and then the audio playback devices whose target characteristic matching degree is greater than the preset characteristic matching degree threshold are determined as the target audio playback devices.
[0062] In one implementation, when predicting behavioral habit characteristics and emotional characteristics, the prediction can be performed through a pre-trained artificial intelligence model to obtain a feature prediction result. When performing matching, the feature matching degree can also be calculated through a preset feature difference algorithm. The feature difference algorithm can be the Pearson Correlation algorithm, the Spearman's Rank Correlation algorithm, the Mutual Information algorithm, etc., which are not specifically limited here.
[0063] The above describes the specific method for each item of target data to determine the target audio playback device. For the case where the number of items of target data is greater than one, the target audio playback devices determined by each item of target data can be fused and calculated to determine the final target audio playback device for playing the target audio.
[0064] When performing the fusion calculation, the number of target audio playback devices determined by each item of target data can be counted, and the audio playback device with a total number greater than a preset total threshold or the largest total number can be determined as the target audio playback device for playing the target audio.
[0065] In one implementation, when determining at least one target audio playback device from an audio playback system according to at least one item of target data, it includes: calculating the matching degree between each item of target data and each audio playback device to obtain a matching degree calculation result corresponding to each item of target data; performing weighted superposition on the matching degree calculation result corresponding to each item of target data according to the preset weight information corresponding to each item of target data to obtain the target matching degree of each audio playback device; and determining the audio playback device with a target matching degree higher than the preset device matching degree threshold as the target audio playback device.
[0066] When calculating the matching degree between each item of target data and each audio playback device, the specific method for each item of target data to determine the target audio playback device described above can be used for calculation to obtain the matching degree calculation result corresponding to each item of target data.
[0067] For example, if the target data is the position information of the detection target, the corresponding matching degree calculation result can be the above-mentioned proximity degree; if the target data is the configuration information of each audio playback device in the audio playback system, the corresponding matching degree calculation result can be the above-mentioned device adaptability; if the target data is the historical control information of the audio playback system, the corresponding matching degree calculation result can be the above-mentioned historical preference degree; if the target data is the attribute information of the target audio, the corresponding matching degree calculation result can be the matching degree of the content type; if the target data is the current spatio-temporal information where the audio playback system is located, the corresponding matching degree calculation result can be the above-mentioned feature matching degree.
[0068] After obtaining the matching degree calculation result, the preset weight information corresponding to each piece of target data can be weighted and superimposed with the corresponding matching degree calculation result to obtain the target matching degree corresponding to each audio playback device, and then the audio playback device with the target matching degree higher than the preset device matching degree threshold can be determined as the target audio playback device, so that the determination of the target audio playback device can be calculated by integrating multi-dimensional data, improving the accuracy of the determination of the playback device and enhancing the intelligence level of the system.
[0069] In one implementation, the preset weight information corresponding to the jth piece of target data can be determined by the following formula:
[0070] where is the preset basic weight of the jth piece of target data, is the learning rate coefficient of the jth piece of target data (belonging to [0, 1]), which is a hyperparameter used to control the degree of adjusting parameters in each step of the optimization algorithm of the machine learning model, is the historical success rate index (belonging to [0, 1]), which is related to whether the matching degree calculation result of the corresponding target data successfully predicts the device actually used by the user for audio playback.
[0071] For example, assume that the 1st (j = 1) piece of target data is the location information of the detection target, the 2nd (j = 2) piece of target data is the configuration information of each audio playback device in the audio playback system, the 3rd (j = 3) piece of target data is the historical control information of the audio playback system, and the 4th (j = 4) piece of target data is the current spatio-temporal information where the audio playback system is located.
[0072] Let the preset basic weights be: = 0.3, = 0.3, = 0.2, = 0.2; the historical success rate index is: = 0.6, = 0.8, = 0.9, = 0.7; the learning rate coefficient is: = 0.5, = 0.5, = 0.7, = 0.6.
[0073] Furthermore, calculating the preset weight information corresponding to the jth piece of target data is: = 0.3 × (1 + 0.5 × 0.6) = 0.39; = 0.3 × (1 + 0.5 × 0.8) = 0.42; = 0.2 × (1 + 0.7 × 0.9) = 0.326; = 0.2 × (1 + 0.6 × 0.7) = 0.284; It is also possible to normalize to obtain: = 0.27, = 0.29, = 0.23, = 0.21; The above algorithm for preset weight information realizes the intelligent fusion of multi-dimensional data through dynamic weight coefficients, can adaptively adjust the influence weights of each dimension according to historical usage conditions, and forms an intelligent system that continuously learns and optimizes.
[0074] After obtaining the dynamic preset weight information, the weighted superposition formula preset below can be used to perform weighted superposition on the calculation results of the matching degrees corresponding to each target data, so as to obtain the target matching degree of the i-th audio playback device :
[0075] Among them, represents the calculation result of the matching degree corresponding to the j-th target data of the i-th audio playback device.
[0076] Suppose there are three smart speakers in the family, and the speaker numbers are 1, 2, and 3 in sequence, located in the living room (i = 1), bedroom (i = 2), and kitchen (i = 3) respectively. For the smart speaker in the living room (i.e., the first smart speaker), the are respectively: = 70%, = 95%, = 60%, = 80%. For the smart speaker in the bedroom (i.e., the second smart speaker), the are respectively: = 85%, = 75%, = 90%, = 85%. For the smart speaker in the kitchen (i.e., the third smart speaker), the are respectively: = 50%, = 70%, = 40%, = 60%.
[0077] Then, the target matching degrees corresponding to these three smart speakers in sequence are as follows: = 0.27×70% + 0.29×95% + 0.23×60% + 0.21×80% = 77.15%; = 0.27×85% + 0.29×75% + 0.23×90% + 0.21×85% = 83.2%; = 0.27×50% + 0.29×70% + 0.23×40% + 0.21×60% = 55.9%; If the preset device matching degree threshold is 70%, then, since = 77.15% > 70%, therefore, the smart speaker in the living room is determined as the target audio playback device. After that, after selecting the actual audio playback device for playback on the terminal, it can be determined whether the target audio playback device is successfully predicted. If the target audio playback device is the same as the actual audio playback device selected by the terminal for playback, it indicates that this prediction is successful; otherwise, it indicates that this prediction fails. The data of this prediction result (success or failure) can be fed back to the learning module to update the historical success rate index and learning rate coefficient of each dimension.
[0078] In one implementation, when playing the target audio through the target audio playback device, it includes: sending the identification information of the target audio playback device to the target terminal; in response to the audio playback device selection made by the terminal based on the target audio playback device, determining the first audio playback device selected by the target terminal for playing the target audio; playing the target audio through the first audio playback device; and adjusting the preset weight information corresponding to each piece of target data according to the first audio playback device.
[0079] Before playing the target audio, the identification information of the determined target audio playback device can also be sent to the target terminal, so that the target terminal can quickly confirm to play the target audio using the target audio playback device, or switch to other audio playback devices to play the target audio. The audio playback device selected by the user through the target terminal for playing the target audio is determined as the first audio playback device.
[0080] The target terminal can be the terminal that triggers the audio playback instruction, or any terminal in the audio playback system, and specific details are not limited here. After the user selects the first audio playback device, the preset weight information corresponding to each piece of target data can be adjusted according to the selected first audio playback device. When weighted superposition is performed on the matching degree calculation results corresponding to each piece of target data according to the adjusted preset weight information, among the target matching degrees of each audio playback device obtained, only the target matching degree of the first audio playback device is higher than the preset device matching degree threshold. This enables the determination of the playback device to adjust the parameters of the algorithm through the user's selection, allowing the user to participate in the training and optimization of the algorithm, making the algorithm increasingly close to the user's expectations.
[0081] Corresponding to the above method embodiments, refer to Figure 2 As shown in the schematic diagram of an audio playback device, which is applied to an audio playback system. The audio playback system includes at least two audio playback devices. The device includes: an acquisition module 20, configured to acquire at least one piece of target data in response to a playback instruction for a target audio; wherein, the at least one piece of target data includes any item of the position information of the detection target, the configuration information of each audio playback device in the audio playback system, the historical control information of the audio playback system, the attribute information of the target audio, and the current spatio-temporal information where the audio playback system is located; a determination module 22, configured to determine at least one target audio playback device from the audio playback system according to the at least one piece of target data; a playback module 24, configured to play the target audio through the target audio playback device.
[0082] The above audio playback device predicts the audio playback device through target data in different dimensions, enabling the intelligent selection of the playback device for audio playback, rather than simply presetting the playback device through default settings, thus improving the intelligence level of the audio playback system.
[0083] Optionally, the at least one piece of target data is the position information of the detection target, and the detection target is a target device or a target organism; the above determination module 22 is configured to: calculate the proximity between the detection target and the preset audible ranges of each audio playback device in the audio playback system according to the position information; determine at least one target audio playback device from the audio playback system whose proximity is greater than or equal to a preset proximity threshold.
[0084] Optionally, the at least one piece of target data is: configuration information of each audio playback device in the audio playback system; the determining module 22 is configured to: obtain the audio-visual type of the target audio; wherein the audio-visual type is used to indicate whether the target audio is a movie and television audio or a non-movie and television audio; calculate the device adaptation degrees of the audio playback devices in the audio playback system according to the audio-visual type and the configuration information of each audio playback device; and determine the audio playback devices with device adaptation degrees greater than a preset device adaptation degree threshold as target audio playback devices, so as to obtain at least one target audio playback device.
[0085] Optionally, the at least one piece of target data is: historical control information of the audio playback system; the determining module 22 is configured to: obtain the current time information and the style label information of the target audio; and match at least one target audio playback device with a historical preference degree higher than a preset preference degree threshold from the historical control information according to the current time information and the style label information.
[0086] Optionally, the at least one piece of target data is: attribute information of the target audio; the determining module 22 is configured to: determine the content type of the target audio according to the attribute information; and determine at least one target audio playback device that matches the preset adaptation content type of the audio playback device from the audio playback system according to the content type.
[0087] Optionally, the at least one piece of target data is: the current spatio-temporal information where the audio playback system is located, and the current spatio-temporal information includes current weather information and / or current time information; the determining module 22 is configured to: predict the behavior habit characteristics and emotional characteristics according to the current spatio-temporal information to obtain a characteristic prediction result; and match at least one target audio playback device with a characteristic matching degree greater than a preset characteristic matching degree threshold with the characteristic prediction result according to the preset characteristic information of each audio playback device in the audio playback system.
[0088] Optionally, the determining module 22 is configured to: calculate the matching degree between each piece of target data and each audio playback device to obtain a matching degree calculation result corresponding to each piece of target data; perform weighted superposition on the matching degree calculation results corresponding to each piece of target data according to the preset weight information corresponding to each piece of target data to obtain the target matching degree of each audio playback device; and determine the audio playback devices with target matching degrees higher than a preset device matching degree threshold as target audio playback devices.
[0089] Optionally, the above-mentioned playback module 24 is specifically configured to: send the identification information of the target audio playback device to the target terminal; in response to the audio playback device selection made by the terminal based on the target audio playback device, determine the first audio playback device selected by the target terminal for playing the target audio; play the target audio through the first audio playback device; and adjust the preset weight information corresponding to each piece of target data according to the first audio playback device.
[0090] This embodiment further provides an electronic device, including a processor and a memory. The memory stores machine-executable instructions that can be executed by the processor, and the processor executes the machine-executable instructions to implement the above-mentioned audio playback method. The electronic device can be a server or a terminal device.
[0091] See Figure 3 As shown, the electronic device includes a processor 100 and a memory 101. The memory 101 stores machine-executable instructions that can be executed by the processor 100, and the processor 100 executes the machine-executable instructions to implement the above-mentioned audio playback method.
[0092] Furthermore, Figure 3 the electronic device shown further includes a bus 102 and a communication interface 103. The processor 100, the communication interface 103, and the memory 101 are connected through the bus 102.
[0093] Among them, the memory 101 may include a high-speed random access memory (RAM, Random Access Memory), and may also include a non-volatile memory, such as at least one disk memory. The communication connection between the system network element and at least one other network element is realized through at least one communication interface 103 (which can be wired or wireless), and the Internet, wide area network, local area network, metropolitan area network, etc. can be used. The bus 102 can be an ISA bus, a PCI bus, an EISA bus, etc. The bus can be divided into an address bus, a data bus, a control bus, etc. For the sake of representation, Figure 3 only a bidirectional arrow is used in [description], but it does not mean that there is only one bus or one type of bus.
[0094] The processor 100 may be an integrated circuit chip with the ability to process signals. In the implementation process, each step of the above method can be completed by the integrated logic circuit in the hardware of the processor 100 or instructions in the form of software. The above-mentioned processor 100 may be a general-purpose processor, including a central processing unit (CPU for short), a network processor (NP for short), etc.; it may also be a digital signal processor (DSP for short), an application specific integrated circuit (ASIC for short), a field-programmable gate array (FPGA for short), or other programmable logic devices, discrete gate or transistor logic devices, discrete hardware components. It can implement or execute the various methods, steps, and logic block diagrams disclosed in the embodiments of the present disclosure. The general-purpose processor may be a microprocessor or the processor may also be any conventional processor, etc. The steps of the method disclosed in combination with the embodiments of the present disclosure can be directly embodied as being executed and completed by a hardware decoding processor, or executed and completed by a combination of hardware and software modules in the decoding processor. The software module may be located in a mature storage medium in the art such as a random access memory, a flash memory, a read-only memory, a programmable read-only memory, or an electrically erasable programmable memory, a register, etc. This storage medium is located in the memory 101, and the processor 100 reads the information in the memory 101 and combines its hardware to complete the steps of the method in the foregoing embodiments, for example: In response to a play instruction for a target audio, obtain at least one piece of target data; wherein, the at least one piece of target data includes: any one of the position information of the detection target, the configuration information of each audio playback device in the audio playback system, the historical control information of the audio playback system, the attribute information of the target audio, and the current spatio-temporal information where the audio playback system is located; determine at least one target audio playback device from the audio playback system according to the at least one piece of target data; play the target audio through the target audio playback device.
[0095] In this way, by predicting the audio playback device through target data in different dimensions, the playback of the audio can intelligently select the playback device, rather than simply presetting the playback device through default settings by intelligence, and the intelligence level of the audio playback system is improved.
[0096] Optionally, the at least one piece of target data is: the location information of the detection target, where the detection target is a target device or a target organism; the step of determining at least one target audio playback device from the audio playback system according to the at least one piece of target data includes: calculating the proximity between the detection target and the preset audible ranges of the audio playback devices in the audio playback system according to the location information; determining at least one target audio playback device in the audio playback system whose proximity is greater than or equal to a preset proximity threshold.
[0097] Optionally, the at least one piece of target data is: the configuration information of the audio playback devices in the audio playback system; the step of determining at least one target audio playback device from the audio playback system according to the at least one piece of target data includes: obtaining the audio-visual type of the target audio; where the audio-visual type is used to indicate that the target audio is a movie and television audio or a non-movie and television audio; calculating the device fitness of the audio playback devices in the audio playback system according to the audio-visual type and the configuration information of each audio playback device; determining the audio playback devices with a device fitness greater than a preset device fitness threshold as target audio playback devices to obtain at least one target audio playback device.
[0098] Optionally, the at least one piece of target data is: the historical control information of the audio playback system; the step of determining at least one target audio playback device from the audio playback system according to the at least one piece of target data includes: obtaining the current time information and the style label information of the target audio; matching at least one target audio playback device with a historical preference degree higher than a preset preference degree threshold from the historical control information according to the current time information and the style label information.
[0099] Optionally, the at least one piece of target data is: the attribute information of the target audio; the step of determining at least one target audio playback device from the audio playback system according to the at least one piece of target data includes: determining the content type of the target audio according to the attribute information; determining at least one target audio playback device in the audio playback system that matches the preset adapted content type according to the content type.
[0100] Optionally, the at least one piece of target data is: the current spatio-temporal information where the audio playback system is located, and the current spatio-temporal information includes current weather information and / or current time information; the step of determining at least one target audio playback device from the audio playback system according to the at least one piece of target data includes: predicting behavioral habit characteristics and emotional characteristics according to the current spatio-temporal information to obtain a characteristic prediction result; matching at least one target audio playback device whose characteristic matching degree with the characteristic prediction result is greater than a preset characteristic matching degree threshold according to the preset characteristic information of each audio playback device in the audio playback system.
[0101] Optionally, the step of determining at least one target audio playback device from the audio playback system according to the at least one piece of target data includes: calculating the matching degree between each piece of target data and each audio playback device to obtain a matching degree calculation result corresponding to each piece of target data; performing weighted superposition on the matching degree calculation result corresponding to each piece of target data according to the preset weight information corresponding to each piece of target data to obtain the target matching degree of each audio playback device; determining the audio playback device with a target matching degree higher than the preset device matching degree threshold as the target audio playback device.
[0102] Optionally, the step of playing the target audio through the target audio playback device includes: sending the identification information of the target audio playback device to the target terminal; in response to the audio playback device selection made by the terminal based on the target audio playback device, determining a first audio playback device selected by the target terminal to play the target audio; playing the target audio through the first audio playback device; adjusting the preset weight information corresponding to each piece of target data according to the first audio playback device.
[0103] This embodiment further provides a computer-readable storage medium, which stores computer-executable instructions. When the computer-executable instructions are called and executed by a processor, the computer-executable instructions cause the processor to implement the above audio playback method, for example: In response to a playback instruction for a target audio, obtaining at least one piece of target data; wherein, the at least one piece of target data includes any item of the position information of the detection target, the configuration information of each audio playback device in the audio playback system, the historical control information of the audio playback system, the attribute information of the target audio, and the current spatio-temporal information where the audio playback system is located; determining at least one target audio playback device from the audio playback system according to the at least one piece of target data; playing the target audio through the target audio playback device.
[0104] In this method, the audio playback device is predicted through target data in different dimensions, enabling the intelligent selection of the playback device for audio playback, rather than simply presetting the playback device through default settings. The intelligence level of the audio playback system is improved.
[0105] Optionally, the at least one piece of target data is: the location information of the detection target, where the detection target is a target device or a target organism; the step of determining at least one target audio playback device from the audio playback system according to the at least one piece of target data includes: calculating the proximity between the detection target and the preset audible ranges of each audio playback device in the audio playback system according to the location information; determining at least one target audio playback device in the audio playback system whose proximity is greater than or equal to a preset proximity threshold.
[0106] Optionally, the at least one piece of target data is: the configuration information of each audio playback device in the audio playback system; the step of determining at least one target audio playback device from the audio playback system according to the at least one piece of target data includes: obtaining the audio-visual type of the target audio; where the audio-visual type is used to indicate whether the target audio is a movie and television audio or a non-movie and television audio; calculating the device adaptation degree of each audio playback device in the audio playback system according to the audio-visual type and the configuration information of each audio playback device; determining the audio playback device with a device adaptation degree greater than a preset device adaptation degree threshold as the target audio playback device to obtain at least one target audio playback device.
[0107] Optionally, the at least one piece of target data is: the historical control information of the audio playback system; the step of determining at least one target audio playback device from the audio playback system according to the at least one piece of target data includes: obtaining the current time information and the style label information of the target audio; matching at least one target audio playback device with a historical preference degree higher than a preset preference degree threshold from the historical control information according to the current time information and the style label information.
[0108] Optionally, the at least one piece of target data is: the attribute information of the target audio; the step of determining at least one target audio playback device from the audio playback system according to the at least one piece of target data includes: determining the content type of the target audio according to the attribute information; determining at least one target audio playback device in the audio playback system that matches the preset adapted content type according to the content type.
[0109] Optionally, the at least one piece of target data is: the current spatio-temporal information where the audio playback system is located, and the current spatio-temporal information includes current weather information and / or current time information; the step of determining at least one target audio playback device from the audio playback system according to the at least one piece of target data includes: predicting behavioral habit characteristics and emotional characteristics according to the current spatio-temporal information to obtain a characteristic prediction result; according to the preset characteristic information of each audio playback device in the audio playback system, matching at least one target audio playback device whose characteristic matching degree with the characteristic prediction result is greater than a preset characteristic matching degree threshold.
[0110] Optionally, the step of determining at least one target audio playback device from the audio playback system according to the at least one piece of target data includes: calculating the matching degree between each piece of target data and each audio playback device to obtain a matching degree calculation result corresponding to each piece of target data; according to the preset weight information corresponding to each piece of target data, performing weighted superposition on the matching degree calculation result corresponding to each piece of target data to obtain the target matching degree of each audio playback device; determining the audio playback device with a target matching degree higher than the preset device matching degree threshold as the target audio playback device.
[0111] Optionally, the step of playing the target audio through the target audio playback device includes: sending the identification information of the target audio playback device to the target terminal; in response to the audio playback device selection made by the terminal based on the target audio playback device, determining the first audio playback device selected by the target terminal to play the target audio; playing the target audio through the first audio playback device; adjusting the preset weight information corresponding to each piece of target data according to the first audio playback device.
[0112] The computer program product of the audio playback method, device, electronic device and storage medium provided by the embodiments of the present disclosure includes a computer-readable storage medium storing program code, and the instructions included in the program code can be used to execute the method described in the foregoing method embodiments. For specific implementation, reference can be made to the method embodiments and will not be elaborated here.
[0113] Those skilled in the art can clearly understand that for the convenience and brevity of description, the specific working processes of the systems and devices described above can refer to the corresponding processes in the foregoing method embodiments and will not be elaborated here.
[0114] In addition, in the description of the embodiments of the present disclosure, unless otherwise clearly specified and limited, the terms "install", "connect", and "couple" should be understood in a broad sense. For example, it may be a fixed connection, a detachable connection, or an integral connection; it may be a mechanical connection or an electrical connection; it may be directly connected or indirectly connected through an intermediate medium, and it may be the communication inside two components. For those skilled in the art, the specific meanings of the above terms in the present disclosure can be understood according to specific situations.
[0115] If the above-mentioned function is implemented in the form of a software functional unit and sold or used as an independent product, it can be stored in a computer-readable storage medium. Based on such an understanding, the technical solution of the present disclosure, in essence, or the part that contributes to the prior art, or a part of this technical solution, can be embodied in the form of a software product. This computer software product is stored in a storage medium and includes several instructions for causing a computer device (which may be a personal computer, a server, or a network device, etc.) to execute all or part of the steps of the methods described in the various embodiments of the present disclosure. The aforementioned storage medium includes: various media such as USB flash drives, mobile hard disks, read-only memories (ROM, Read-Only Memory), random access memories (RAM, Random Access Memory), magnetic disks, or optical discs that can store program codes.
[0116] In the description of the present disclosure, it should be noted that the orientation or positional relationship indicated by the terms "center", "upper", "lower", "left", "right", "vertical", "horizontal", "inner", "outer", etc. is based on the orientation or positional relationship shown in the drawings. It is only for the convenience of describing the present disclosure and simplifying the description, rather than indicating or implying that the device or element referred to must have a specific orientation, be constructed and operated in a specific orientation. Therefore, it should not be construed as a limitation to the present disclosure. In addition, the terms "first", "second", and "third" are only used for descriptive purposes and cannot be understood as indicating or implying relative importance.
[0117] Finally, it should be noted that the above embodiments are only specific implementation manners of the present disclosure, used to illustrate the technical solutions of the present disclosure, rather than limiting them. The protection scope of the present disclosure is not limited thereto. Although the present disclosure has been described in detail with reference to the foregoing embodiments, those skilled in the art should understand that: any person skilled in the art within the technical scope disclosed by the present disclosure can still modify the technical solutions recorded in the foregoing embodiments, or can easily think of changes, or perform equivalent replacements on some of the technical features; and these modifications, changes, or replacements do not make the essence of the corresponding technical solutions deviate from the spirit and scope of the technical solutions of the embodiments of the present disclosure, and should all be covered by the protection scope of the present disclosure. Therefore, the protection scope of the present disclosure should be subject to the protection scope of the claims.
Claims
1. A method for playing audio, characterized in that: Applied to an audio playback system, the audio playback system includes at least two audio playback devices, and the method includes: In response to a play instruction for a target audio, at least one target data is acquired; wherein the at least one target data includes: any item of location information of a detection target, configuration information of each audio playback device in the audio playback system, historical control information of the audio playback system, attribute information of the target audio, and current space-time information of the audio playback system; Determining at least one target audio playback device from the audio playback system according to the at least one item of target data; The target audio is played through the target audio playing device.
2. The method according to claim 1, characterized in that The at least one target data is: location information of a detection target, the detection target being a target device or a target organism; The step of determining at least one target audio playback device from the audio playback system according to the at least one target data comprises: Calculating the proximity between the detection target and the preset audible range of each audio playback device in the audio playback system according to the position information; At least one target audio playback device whose proximity is greater than or equal to a preset proximity threshold is determined from the audio playback system.
3. The method according to claim 1, characterized in that: The at least one target data is: configuration information of each audio playback device in the audio playback system; The step of determining at least one target audio playback device from the audio playback system according to the at least one target data comprises: Acquire the audio / video type of the target audio; wherein the audio / video type is used to indicate whether the target audio is video audio or non-video audio; Calculating the device adaptability of each audio playback device in the audio playback system according to the audio and video type and the configuration information of each audio playback device; The audio playback device whose device adaptability is greater than a preset device adaptability threshold is determined as a target audio playback device, and at least one target audio playback device is obtained.
4. The method according to claim 1, characterized in that: The at least one target data is: historical control information of the audio playback system; The step of determining at least one target audio playback device from the audio playback system according to the at least one target data comprises: Obtaining current time information and style tag information of the target audio; According to the current time information and the style tag information, at least one target audio playback device having a history preference higher than a preset preference threshold is matched from the history control information.
5. The method according to claim 1, characterized in that The at least one target data is: attribute information of the target audio; The step of determining at least one target audio playback device from the audio playback system according to the at least one target data comprises: Determining the content type of the target audio according to the attribute information; According to the content type, at least one target audio playback device matching a preset adapted content type of the audio playback device is determined from the audio playback system.
6. The method according to claim 1, characterized in that The at least one target data is: current time and space information of the audio playback system, the current time and space information including current weather information and / or current time information; The step of determining at least one target audio playback device from the audio playback system according to the at least one target data comprises: Predicting behavioral habit characteristics and emotional characteristics based on the current spatiotemporal information to obtain characteristic prediction results; According to the preset feature information of each audio playback device in the audio playback system, at least one target audio playback device having a feature matching degree greater than a preset feature matching degree threshold value is matched with the feature prediction result.
7. The method according to claim 1, characterized in that The step of determining at least one target audio playback device from the audio playback system according to the at least one target data comprises: Calculate the matching degree between each target data and each audio playback device to obtain the matching degree calculation result corresponding to each target data; According to the preset weight information corresponding to each item of target data, the matching degree calculation results corresponding to each item of target data are weighted and superimposed to obtain the target matching degree of each audio playback device; An audio playback device whose target matching degree is higher than a preset device matching degree threshold is determined as a target audio playback device.
8. The method according to claim 7, characterized in that The step of playing the target audio through the target audio playing device includes: Sending identification information of the target audio playback device to a target terminal; In response to the audio playback device selection performed by the terminal based on the target audio playback device, determining that the target terminal selects a first audio playback device for playing the target audio; Playing the target audio through the first audio playback device; According to the first audio playback device, preset weight information corresponding to each item of target data is adjusted.
9. An audio playback device, characterized in that: Applied to an audio playback system, the audio playback system includes at least two audio playback devices, and the apparatus includes: An acquisition module, configured to acquire at least one target data in response to a playback instruction for a target audio; wherein the at least one target data includes: any item of location information of a detection target, configuration information of each audio playback device in the audio playback system, historical control information of the audio playback system, attribute information of the target audio, and current space-time information of the audio playback system; A determination module, configured to determine at least one target audio playback device from the audio playback system according to the at least one target data; The playing module is used to play the target audio through the target audio playing device.
10. An electronic device, characterized in that: It comprises a processor and a memory, wherein the memory stores machine executable instructions that can be executed by the processor, and the processor executes the machine executable instructions to implement the audio playing method according to any one of claims 1-8.
11. A computer-readable storage medium, characterized in that: The computer-readable storage medium stores computer-executable instructions. When the computer-executable instructions are called and executed by a processor, the computer-executable instructions prompt the processor to implement the audio playback method described in any one of claims 1-8.
Citation Information
Patent Citations
Playing equipment selection method, device and equipment and computer readable storage medium
CN112333533A
Audio playing method and device, electronic equipment and storage medium
CN113296728A
Playing equipment control method and device, equipment and medium
CN115550755A
Bluetooth sound box control method and system and Bluetooth sound box
CN117119352A
Audio playing method and electronic equipment
CN117992008A